Audiomatic is a web app that seamlessly translates videos into other languages. Our state-of-the-art pipeline delivers contextually-accurate dubbed translations that preserve the tone, style, and emotion of the original speakers.



AudiowaveAI is described as 'Lets you convert text to high-quality audio easily and affordably. Listen to PDFs, epubs, articles, blog posts, links, emails or anything else you want on any device with natural-sounding voices' and is a Text to Speech service. There are more than 25 alternatives to AudiowaveAI for a variety of platforms, including Web-based, Windows, Mac, Linux and Android apps. The best AudiowaveAI alternative is ElevenLabs, which is free. Other great apps like AudiowaveAI are VoiceCraft, SherpaTTS , X to Voice and NaturalReader.
Audiomatic is a web app that seamlessly translates videos into other languages. Our state-of-the-art pipeline delivers contextually-accurate dubbed translations that preserve the tone, style, and emotion of the original speakers.



We're excited to introduce Chatterbox, Resemble AI's first production-grade open source TTS model. Licensed under MIT, Chatterbox has been benchmarked against leading closed-source systems like ElevenLabs, and is consistently preferred in side-by-side evaluations.
AI-powered generator for Windows and macOS offering instant offline voice cloning from 3-second samples, thousands of voices, over 600 languages, unlimited generations, long-form audiobook creation, transcription, multi-voice casting, and commercial use rights.




The eSpeak NG is a compact open source software text-to-speech synthesizer for Linux, Windows, Android and other operating systems. It supports more than 100 languages and accents. It is based on the eSpeak engine created by Jonathan Duddington.
TTSMaker is a free text-to-speech tool that provides speech synthesis services, supports multiple languages: English, French, German, Spanish, Arabic, Chinese, Japanese, Korean, Vietnamese... and a variety of voice styles, you can use it reads text and e-books aloud, and can...

Supertonic TTS is a high-performance, on-device text-to-speech (TTS) system for Android. It uses Supertonic ONNX models to provide high-quality speech synthesis entirely offline, ensuring privacy and availability without an internet connection.



Free open source AI voice cloning and text to speech synthesis. Clone a voice in 5 seconds to generate arbitrary speech in real-time.
Processes both audio and text inputs with a multi-task framework supporting over 30 language and sound tasks, enabling multi-turn dialogue, sound reasoning, and tool use, while excelling in benchmarks without task-specific fine-tuning or retraining.


QwenVoice is a native SwiftUI macOS application that brings state-of-the-art text-to-speech to Apple Silicon Macs with no Python install, no terminal, and no dependencies required of the user — just download and run.



Dia is a 1.6B parameter text to speech model created by Nari Labs. It was pushed to the Hub using the PytorchModelHubMixin integration.

Sick of subscriptions? We got you. ARES lets you use multiple AI tools without having to get a subscription.

Mimic is a powerful TTS tool. Mimic is low-latency and has a small resource footprint. Its range of high quality voices also set it apart from other open source text-to-speech projects.