Handy
Handy is a free, open source, cross-platform speech-to-text app that runs completely offline and puts whatever you say directly into any text field. Press and hold a configurable keyboard shortcut, speak, and release; Handy records your voice, transcribes it locally, and pastes the result into the app you are using. Push-to-talk is enabled by default, but users can switch to a toggle mode that starts and stops recording with separate key presses. Handy supports macOS, Windows, and Linux and keeps voice data on the computer instead of sending audio to cloud services. Users can choose between Whisper models and Parakeet V3, Whisper provides broad multilingual support across more than 99 languages, while Parakeet V3 is optimized for fast CPU performance and automatic language detection. Silence is filtered with voice activity detection, and GPU acceleration is available for Whisper where supported.
Learn more
Superwhisper
Superwhisper is a voice-to-text app that lets users dictate, transcribe, and control writing workflows across apps without relying on the keyboard. The platform works on Mac, Windows, and iOS, with voice input that can be used in tools such as Slack, Cursor, Notion, Claude Code, Codex, and other agentic coding apps. Superwhisper supports push-to-talk, custom shortcuts, file transcription, meeting recording, custom modes, vocabulary settings, and AI-enhanced output. Users can create modes for different tasks, languages, tones, formats, prompts, and applications. The platform supports more than 100 languages and lets users choose from models such as GPT, Claude, Llama, Grok, Gemini, and others. Built for fast-moving professionals, developers, writers, and teams, Superwhisper helps turn speech into polished text, commands, transcripts, and AI-ready prompts.
Learn more
SpokenData
Let the automatic speech-to-text technology transcribe your data. Or transcribe your data yourself or buy professional transcript. Use our on-line time synchonous editor to surf your data and transcripts. Download transcripts in many formats. Manage your team of transcribers using tags and categories. Help them with transcription by automatic voice-to-text technology. Integrate SpokenData into your application via our REST API. We adapt the voice-to-text on your data domain to maximize the transcript accuracy and lower your labor costs. Enable speech technologies in your applications through integrating SpokenData using our REST API. We are ready to process huge amounts of your data. You get API fitting your needs. Just contact our support team. We customize the voice-to-text on your data and purpose to maximize the transcript accuracy. Suitable for: web/mobile app developers, media monitoring agencies, audio/video archive business.
Learn more
VoiceTypr
VoiceTypr is an offline, AI-powered voice-to-text tool available for both Windows and macOS that lets you dictate anywhere you can type by simply holding or toggling a hotkey, with automatic transcription directly into applications such as chat editors, code editors, email fields, and text boxes. It supports over 100 languages, offers multiple transcription-model choices (focusing on accuracy or speed), includes smart formatting modes for everything from casual chat to formal documents, and maintains a searchable history of transcriptions that you can export or copy. Crucially, all processing occurs locally on your machine, so your audio stays private. You simply install the app, download your preferred model, set a global hotkey, then speak and ship, whether you’re writing code prompts, emails, notes, or messages. Additional features include drag-and-drop transcription of MP3, WAV, M4A, MP4, or MOV files, global hotkey activation, and hardware hardware-accelerated performance.
Learn more