Instant Chrome extension that translates Udemy courses while you watch. Listen to lectures in your language with natural AI voice instead of reading subtitles all the time.
Cost / License
- Freemium
- Proprietary
Platforms
- Google Chrome
Estonia
EU



Apps similar to Kokoro include VoiceStudio, which is free and open source. Other options are ElevenLabs, SherpaTTS , ElevenReader and Balabolka. Text to Speech Services like Kokoro are a crowded field, with more than 100 alternatives for Windows, the web, Mac, Android and iPhone.
Instant Chrome extension that translates Udemy courses while you watch. Listen to lectures in your language with natural AI voice instead of reading subtitles all the time.




NextUp.com develops Windows text to speech (TTS) software applications like TextAloud that let your computer talk with AT&T Natural Voices. TextAloud can also be in Microsoft Word as a plug-in.

We're excited to introduce Chatterbox, Resemble AI's first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations.
CloudTTS is a straightforward text-to-speech application. Simply type in or paste the text you'd like to hear, and it reads it back to you.

Tontaube is a Text-To-Speech & Audiobook streaming platform that offers a library of over 30,000 literary classics, available to read or listen to. We continuously convert these texts into audiobooks with the best AI voices on the market.



Transforms digital and printed content into natural-sounding speech across devices, with adjustable speed, offline playback, multiple voice options, support for various file formats, and mobile or web app access, including features for accessibility needs.




AIVocal is your all-in-one AI assistant for voice tasks—perfect for AI podcasting, speech generation, vocal editing, and voice control. From transcribing meetings to creating high-quality audio content, AIVocal makes voice work smarter and faster.

The eSpeak NG is a compact open source software text-to-speech synthesizer for Linux, Windows, Android and other operating systems. It supports more than 100 languages and accents. It is based on the eSpeak engine created by Jonathan Duddington.
AI voice platform features 60+ emotional voices in multiple languages and accents for commercial-grade text-to-speech, supports voice cloning for personal use, offers APIs for workflow integration, enables digital preservation, and fits various audio projects.


TTSMaker is a free text-to-speech tool that provides speech synthesis services, supports multiple languages: English, French, German, Spanish, Arabic, Chinese, Japanese, Korean, Vietnamese... and a variety of voice styles, you can use it reads text and e-books aloud, and can...

Voicebox is a state-of-the-art speech generative model built upon Meta’s non-autoregressive flow matching model. By learning to solve a text-guided speech infilling task with a large scale of data, Voicebox outperforms single purpose AI models across speech tasks through...
