FluidVoice
FluidVoice is a free, open source macOS dictation app that pairs local speech recognition with Fluid-1, an on-device AI model for polishing dictation. One hotkey lets users speak into any text field across email, documents, chat, terminals, code editors, and other apps, with text appearing in near real time. Local speech models keep dictation on-device and work offline, while optional AI post-processing can use Fluid Intelligence, OpenAI, Groq, or custom providers. Fluid-1 cleans up rough dictation, fixes formatting, capitalization, dates, names, and numbers, and adapts tone to the active app without changing the speaker’s meaning. Users can create custom prompts for different apps, while Write Mode, Command Mode, and Direct Dictation make it easy to switch contexts. FluidVoice supports more than 40 languages across models including Nemotron Speech 3.5, Parakeet Flash, Parakeet TDT v2 and v3, Cohere Transcribe, Apple Speech, and Whisper.
Learn more
DictaFlow
DictaFlow is an AI dictation app for Windows, Mac, iPhone, and Android via Telegram that turns messy speech into clean text wherever the cursor is. Hold a keyboard shortcut, mouse button, or VDI-safe trigger, speak naturally, and release to insert text directly into email, documents, IDEs, EHRs, browsers, terminals, notes, and remote desktops. It is built for the messy parts of dictation, including names, acronyms, code terms, drug names, clinical shorthand, contract language, accents, and more than 100 languages. DictaFlow understands mid-sentence corrections such as “actually” and “I mean,” so users can fix mistakes without breaking rhythm, and AI cleanup can turn rough speech into emails, bullets, code comments, meeting notes, prompts, or formatted text as they talk. Users can also highlight existing text in apps such as Word, Slack, or VS Code and edit it by voice.
Learn more
Spokenly
Spokenly is an AI-powered dictation app for Mac, iPhone, Windows, and Linux that turns speech into clean, punctuated text wherever you work. Hold a shortcut, speak naturally, and release to place the transcription directly at the cursor in browsers, email, chat, word processors, IDEs, terminals, and other apps. It supports more than 100 languages, including mixed-language dictation, and offers both local and cloud speech-to-text models. Whisper, Parakeet, and other on-device models can run completely offline, while cloud engines from providers such as OpenAI, Deepgram, Groq, Soniox, and ElevenLabs can be used for higher-accuracy or real-time transcription. Local Only Mode blocks network requests so voice data stays on the device. Modes let users save different transcription models, AI providers, prompts, and output styles for specific tasks, while AI Instructions can remove filler words, fix grammar and punctuation, summarize, rewrite, translate, or reformat dictated text.
Learn more
Handy
Handy is a free, open source, cross-platform speech-to-text app that runs completely offline and puts whatever you say directly into any text field. Press and hold a configurable keyboard shortcut, speak, and release; Handy records your voice, transcribes it locally, and pastes the result into the app you are using. Push-to-talk is enabled by default, but users can switch to a toggle mode that starts and stops recording with separate key presses. Handy supports macOS, Windows, and Linux and keeps voice data on the computer instead of sending audio to cloud services. Users can choose between Whisper models and Parakeet V3, Whisper provides broad multilingual support across more than 99 languages, while Parakeet V3 is optimized for fast CPU performance and automatic language detection. Silence is filtered with voice activity detection, and GPU acceleration is available for Whisper where supported.
Learn more