Voice-first AI for meeting notes, voice notes, and dictation. 5× faster than typing. Just speak, and it's done.

Voxtral is described as 'State-of-the-art speech models with transcription, translation, and audio understanding, available via API or self-hosted, optimized for cost and efficiency' and is a audio transcription tool in the ai tools & services category. There are more than 100 alternatives to Voxtral for a variety of platforms, including Mac, Web-based, Windows, iPhone and Linux apps. The best Voxtral alternative is Handy STT, which is both free and Open Source. Other great apps like Voxtral are Vibe Transcribe, FUTO Voice Input, TypeWhisper and Spokenly.
Voice-first AI for meeting notes, voice notes, and dictation. 5× faster than typing. Just speak, and it's done.

CMU Sphinx is a speaker-independent large vocabulary continuous speech recognizer released under BSD style license. It is also a collection of open source tools and resources that allows researchers and developers to build speech recognition systems.
Windows Speech Recognition makes using a keyboard and mouse optional. You can control your PC with your voice and dictate text instead.
Amphion is a toolkit for Audio, Music, and Speech Generation. Its purpose is to support reproducible research and help junior researchers and engineers get started in the field of audio, music, and speech generation research and development.
High-quality on-device transcription. Easily convert speech to text from meetings, lectures, and more.

Write with your voice in any app on macOS. Faster and more accurate than ChatGPT, Google and OpenAI Whisper. Start talking. Stop typing.

Convert your audio and video to accurate text in seconds with advanced speaker recognition, and let AI automatically generate notes to quickly uncover the insights you need.



Vocol is an AI transcription software and a one-stop voice collaboration platform designed to boost work efficiency by turning voice and data into actionable insights.



AudioNotes app allows you to effortlessly record, transcribe, and enhance audio from anywhere using AI. Whether you're capturing thoughts, ideas, interviews, meetings, or lectures, this app has you covered.




Voice2Sub is a local-first Whisper AI desktop application for converting audio and video files into subtitles, transcripts, and editable subtitle text. It provides a private speech-to-text workflow for users who want to generate, review, edit, and export subtitle files without...




Glasscribe is a lightweight macOS menu bar app that transcribes speech in real time — entirely on your device. Built on Apple's native Speech framework (macOS 26 Tahoe), it captures both system audio and microphone input across 22+ languages with real-time on-device...