Instant Chrome extension that translates Udemy courses while you watch. Listen to lectures in your language with natural AI voice instead of reading subtitles all the time.
Cost / License
- Freemium
- Proprietary
Platforms
- Google Chrome



SpeakPerfect is described as 'Turns your fuzzy thoughts into great script + audio using AI. With SpeakPerfect, you no longer need to spend hours writing down the script before making a video' and is a Text to Speech service in the ai tools & services category. There are more than 25 alternatives to SpeakPerfect for a variety of platforms, including Web-based, Mac, Windows, Linux and Android apps. The best SpeakPerfect alternative is ElevenLabs, which is free. Other great apps like SpeakPerfect are VoiceCraft, SherpaTTS , X to Voice and NaturalReader.
Instant Chrome extension that translates Udemy courses while you watch. Listen to lectures in your language with natural AI voice instead of reading subtitles all the time.



Supertonic TTS is a high-performance, on-device text-to-speech (TTS) system for Android. It uses Supertonic ONNX models to provide high-quality speech synthesis entirely offline, ensuring privacy and availability without an internet connection.




Free open source AI voice cloning and text to speech synthesis. Clone a voice in 5 seconds to generate arbitrary speech in real-time.
Processes both audio and text inputs with a multi-task framework supporting over 30 language and sound tasks, enabling multi-turn dialogue, sound reasoning, and tool use, while excelling in benchmarks without task-specific fine-tuning or retraining.


QwenVoice is a native SwiftUI macOS application that brings state-of-the-art text-to-speech to Apple Silicon Macs with no Python install, no terminal, and no dependencies required of the user — just download and run.



Dia is a 1.6B parameter text to speech model created by Nari Labs. It was pushed to the Hub using the PytorchModelHubMixin integration.

Creates natural-sounding speech from text in multiple languages and voices using neural networks, with an advanced browser editor, export options as MP3, WAV, or MP4, SSML formatting, transcription, video recording, customization features, and collaboration tools.







Free and offline Text-to-Speech (TTS) engine that reads any text on your screen with high-quality voices powered by AI models.
An AI-powered platform revolutionizing voice creation with cutting-edge technology. We provide advanced audio solutions for creators and businesses worldwide.



Runs chat, image generation, speech synthesis and voice cloning locally on Windows, macOS and Linux. Its bundled inference stack works offline by default, without an account, cloud connection, subscription or external API. USB mode runs without writing to its host.




Convert text into natural-sounding speech using an API powered by the best of Google’s AI technologies.
