VoiceInk
VoiceInk is a native macOS dictation app that uses local AI models to instantly turn speech into clean text with near-perfect accuracy and complete privacy. It works across apps, letting users dictate emails, messages, notes, prompts, documents, and code without changing their workflow. Local models keep audio and processing on the Mac, while cloud providers are optional and only used when connected and selected by the user. Global shortcuts support toggle recording, push-to-talk, retry, cancel, and paste actions without reaching for the app. A personal dictionary teaches VoiceInk names, technical terms, unusual spellings, phrases, and Smart Replace shortcuts for frequently used text. Contextual awareness can use selected text, clipboard content, or visible screen text to improve transcription and AI-enhanced output. Modes let users save different transcription models, enhancement prompts, context settings, output behavior, and shortcuts for specific apps, websites, or tasks.
Learn more
Dictation Speech to Text
You can now add custom words to improve speech recognition! Find the list in setup->manage custom words. Dictation Speech to text allows to dictate, record, translate and transcribe text instead of typing. It uses latest speech to text voice recognition technology and its main purpose is speech to text and translation for text messaging. Never type any text, just dictate and translate using your speech! Nearly every app that can send text messages can be configured to operate with 'Dictation Speech to text'. Dictate uses the builtin speech to text recognition engine. Dictation Speech to text supports more than 40 languages. Dictate offers 3 text zones, indicated by language flags, for which you can configure a different language in the settings. Thus you can switch between different language projects with a singe click. Translation is as easy as pushing the translation button. You can specify the translation target language in the app settings.
Learn more
Amazon Transcribe
Amazon Transcribe makes it easy for developers to add speech to text capabilities to their applications. Audio data is virtually impossible for computers to search and analyze. Therefore, recorded speech needs to be converted to text before it can be used in applications. Historically, customers had to work with transcription providers that required them to sign expensive contracts and were hard to integrate into their technology stacks to accomplish this task. Many of these providers use outdated technology that does not adapt well to different scenarios, like low-fidelity phone audio common in contact centers, which results in poor accuracy. Amazon Transcribe uses a deep learning process called automatic speech recognition (ASR) to convert speech to text quickly and accurately. Amazon Transcribe can be used to transcribe customer service calls, automate subtitling, and generate metadata for media assets to create a fully searchable archive.
Learn more
Wispr Flow
Wispr Flow is an AI voice-to-text app that turns natural speech into clear, polished writing across apps and devices. The platform lets users dictate messages, documents, code, notes, emails, and replies up to four times faster than typing. Wispr Flow automatically removes filler words, fixes typos, improves formatting, and transforms rambling speech into clean written text. It works across Mac, Windows, iPhone, and Android, helping users write wherever they work. The platform includes AI auto-edits, a personal dictionary, snippet shortcuts, multilingual transcription, and support for more than 100 languages. Built for professionals, creators, students, developers, accessibility users, and teams, Wispr Flow helps people capture ideas faster and write more naturally with their voice.
Learn more