Hold a key, speak, release — AI voice-to-text dictation that types into any Windows app. Free & open-source.
Cost / License
- Free
- Open Source (GPL-3.0)
Platforms
- Windows



SpeechPulse is described as 'Dictation software for Windows 10/11 and Apple Silicon Macs. It can type into any text input, including text editors, web browsers, and office applications. SpeechPulse works fully offline and doesn’t require any internet connectivity' and is a audio transcription tool in the audio & music category. There are more than 100 alternatives to SpeechPulse for a variety of platforms, including Mac, Web-based, Windows, iPhone and iPad apps. The best SpeechPulse alternative is Handy STT, which is both free and Open Source. Other great apps like SpeechPulse are Vibe Transcribe, Voxtral, FUTO Voice Input and TypeWhisper.
Hold a key, speak, release — AI voice-to-text dictation that types into any Windows app. Free & open-source.



Hello Transcribe is a private and secure speech to text transcriber that uses OpenAI Whisper and Whisper.cpp.




Meeting Recorder is your personal assistant for meetings. It listens and transcribes meetings and conferences for you, allowing you to search for words and phrases within your recording. You can record your most important conversations and save time, helping you work more...



Vocatim turns recorded or imported audio into editable transcripts on iPhone, iPad and Mac. It supports speaker labels and local AI summaries.



Cadence is a native voice workspace. Because our AI model runs on-device, your voice never leaves your Mac.




HoldToType turns speech into text at the cursor in any Windows program: hold two keys, speak, let go, and the words land where you were typing. Recognition runs on your own computer with open models, and you are not tied to Whisper: Nemotron 3.




Open-source Mac application supporting local transcription of microphone, media, and system audio sources, with editable transcripts, live captions, privacy by design, export in TXT, Markdown, JSON, PDF, SRT, WebVTT, translation, and offline processing. Requires Apple silicon.


FLUENT is a hotkey-activated speech-to-text recognition tool that conveniently displays the recognition results & copies them to the clipboard.




VibeVoice is a novel framework designed for generating expressive, long-form, multi-speaker conversational audio, such as podcasts, from text. It addresses significant challenges in traditional Text-to-Speech (TTS) systems, particularly in scalability, speaker consistency, and...


Transcribe audio and video files in a blink, automatically, all offline, and with highly accurate results. AI Transcription uses OpenAI’s Whisper technology and Apple Speech Recognition to convert speech (like in podcasts, presentations, lectures, or voice messages) into text...




Real-time AI translation for Google Meet enables on-screen live subtitles in 30+ languages with private device storage, zero setup, local record review, search, customizable exports, no bots or interruptions, AI summaries, and seamless instant use.




Transcribe any audio or video to text for free, with automatic speaker labels and no daily cap. You get a TXT transcript plus SRT/VTT subtitles; an hour of audio takes about two minutes.