Hold a key, speak, release — AI voice-to-text dictation that types into any Windows app. Free & open-source.
Cost / License
- Free
- Open Source (GPL-3.0)
Platforms
- Windows



Apps similar to SpeechPulse include Handy STT, which is free and open source. Other options are Vibe Transcribe, Voxtral, FUTO Voice Input and TypeWhisper. You're spoiled for choice with Audio Transcription Tools like SpeechPulse: we list more than 100 alternatives for Mac, the web, Windows, iPhone and iPad.
Hold a key, speak, release — AI voice-to-text dictation that types into any Windows app. Free & open-source.



Hello Transcribe is a private and secure speech to text transcriber that uses OpenAI Whisper and Whisper.cpp.




Meeting Recorder is your personal assistant for meetings. It listens and transcribes meetings and conferences for you, allowing you to search for words and phrases within your recording. You can record your most important conversations and save time, helping you work more...



Vocatim turns recorded or imported audio into editable transcripts on iPhone, iPad and Mac. It supports speaker labels and local AI summaries.



Cadence is a native voice workspace. Because our AI model runs on-device, your voice never leaves your Mac.




SpeechPilot is a Chrome extension for people who type all day in browser-based chat and support tools. Dictate straight into any text field with spoken punctuation in 13 languages (speech recognition in 69 languages).




HoldToType turns speech into text at the cursor in any Windows program: hold two keys, speak, let go, and the words land where you were typing. Recognition runs on your own computer with open models, and you are not tied to Whisper: Nemotron 3.




Open-source Mac application supporting local transcription of microphone, media, and system audio sources, with editable transcripts, live captions, privacy by design, export in TXT, Markdown, JSON, PDF, SRT, WebVTT, translation, and offline processing. Requires Apple silicon.


FLUENT is a hotkey-activated speech-to-text recognition tool that conveniently displays the recognition results & copies them to the clipboard.




VibeVoice is a novel framework designed for generating expressive, long-form, multi-speaker conversational audio, such as podcasts, from text. It addresses significant challenges in traditional Text-to-Speech (TTS) systems, particularly in scalability, speaker consistency, and...


Transcribe audio and video files in a blink, automatically, all offline, and with highly accurate results. AI Transcription uses OpenAI’s Whisper technology and Apple Speech Recognition to convert speech (like in podcasts, presentations, lectures, or voice messages) into text...




Keeps each recording byte-for-byte, transcribes it through a service you choose, and lets you search inside everything you have.



