Open-source AI voice typing for macOS, Windows, and Linux. Press a hotkey, speak naturally, get polished text in any app.
Cost / License
- Freemium
- Open Source (MIT)
Platforms
- Mac
- Windows
- Linux



AI Audio Kit is described as 'A straightforward macOS application that allows the user to use different Whisper services (OpenAI API, Runpod Faster Whisper) from your macOS desktop. You have the flexibility to use your own API key, ensuring that you only incur charges for the services you actively use' and is a audio transcription tool in the ai tools & services category. There are more than 100 alternatives to AI Audio Kit for a variety of platforms, including Mac, Web-based, Windows, iPhone and Android apps. The best AI Audio Kit alternative is Handy STT, which is both free and Open Source. Other great apps like AI Audio Kit are Vibe Transcribe, Voxtral, FUTO Voice Input and TypeWhisper.
Open-source AI voice typing for macOS, Windows, and Linux. Press a hotkey, speak naturally, get polished text in any app.



Transcribe audio and video files in a blink, automatically, all offline, and with highly accurate results. AI Transcription uses OpenAI’s Whisper technology and Apple Speech Recognition to convert speech (like in podcasts, presentations, lectures, or voice messages) into text...




Bulbul is a free, open-source voice dictation app for Windows, macOS, Linux, and Android. Hold a hotkey (or tap the bubble on your phone), speak, and cleaned-up text appears in whatever app you're using — editor, browser, chat, terminal.
In-person conversation recorder with voice-based speaker memory, per-speaker consent management, and cross-recording AI for iPhone, iPad, and Mac.



Hosted EU-based platform for interview transcription with correctable speaker-labeled transcripts, centralized searchable archive, tagging, export to REFI-QDA and major research tools, GDPR compliance, permanent free tier access, and metered audio hour plans.



Wisprs is AI transcription software for audio and video — made easy and fast. Upload a file (client calls, interviews, podcast episodes, voice memos) and get speech-to-text you can edit, with excellent accuracy on clear audio.



Ebby will automatically convert your audio to text for a fraction of the time and cost of traditional services.




AI legal transcription at $0.25/minute with speaker diarization and court-ready formatting. Upload depositions, interviews, and hearings — get accurate, speaker-identified transcripts in minutes. HIPAA compliant. Built for law firms that need admissible transcripts.





Record meetings, lectures, and podcasts. Transcribe in 10+ languages with on-device Apple models. Get ChatGPT-powered summaries via Apple Intelligence — no subscriptions.




WordWand is a system-wide AI assistant for macOS that works in any app through a single keyboard shortcut. No copy-pasting, no tab switching — just select text, press a hotkey, and transform it instantly.



AssemblyAI is API for speech recognition. They’ve built “accurate, simple and customizable” technology that the team claims is what “Stripe did to payments,” but for speech. The voice technology industry is growing fast, due to the popularity of Siri, Alexa and Google Home.
