Free, on-device AI dictation for macOS. Hold a hotkey, speak, paste polished text. No account, no cloud, no subscription.
Cost / License
- Free
- Open Source (GPL-3.0)
Application type
Platforms
- Mac
United States




Whisper is described as 'End-to-end speech recognition model trained on 680,000 hours of multitask, multilingual audio data, offering robust transcription, translation, and language identification' and is a audio transcription tool in the audio & music category. There are more than 100 alternatives to Whisper for a variety of platforms, including Mac, Windows, Web-based, iPhone and Android apps. The best Whisper alternative is Handy STT, which is both free and Open Source. Other great apps like Whisper are Vibe Transcribe, Voxtral, FUTO Voice Input and TypeWhisper.
Free, on-device AI dictation for macOS. Hold a hotkey, speak, paste polished text. No account, no cloud, no subscription.




Power your apps with world-class speech-to-text and domain-specific language models (DSLMs). Effortlessly accurate. Blazing fast. Enterprise-ready scale. Unbeatable pricing. Everything developers need to build with confidence and ship faster.

Gladia is a production-ready Speech-to-Text API built for teams shipping real-world voice products—delivering high accuracy, multilingual coverage, real-time + async transcription, and a growing set of add-ons (diarization, translation, summarization, sentiment, formatting, and more).



Voice-to-text solution for Mac enabling dictation anywhere with a keyboard shortcut—no internet required. Delivers instant, private transcription in 100+ languages, translation to English, vocabulary customization, universal app compatibility, one-time payment, and zero data collection.



Provides instant, local voice-to-text transcription, searchable timeline for both dictated and clipboard text, fully private processing on macOS, adaptive learning for terminology, invisible activation, audio/video file transcription, and seamless pasting into apps.


Converts audio into editable text via AI, supports recording, translation, and summarization. Perfect for meetings with real-time transcription and tool integrations.




Automated platform offering quick, accurate audio and video transcription into editable text or subtitles in over 30 languages, with options for smart punctuation, file uploads from cloud storage services, secure payment, free trial, and team collaboration.



This software translates audio and video into text in over 35 languages, offering an in-browser editor for seamless transcription management. With automated subtitles, language conversion, and media player sharing, it supports collaboration and assures secure data storage with Zoom and Adobe integration.



Transcribe any audio and get fast and accurate transcripts with timestamps using AI. Generate new content from the transcripts such as summaries, blog-posts, social media posts or your own custom content with GPT prompts. No subscription required.


Dictly turns speech into polished, structured text — instantly and entirely on your device. No servers. No data collection. No delay.




AI voice dictation for macOS and iOS adapts to vocabulary and writing style, converts speech into seamless text for any app, syncs preferences and dictionaries, offers multiple dictation modes, and ensures all data is kept private with zero server storage.




Whisperian is a highly configurable voice-to-text (voice typing) tool made for Android. Supports local STT models like Parakeet v3.


