Dictation Speech to Text
You can now add custom words to improve speech recognition! Find the list in setup->manage custom words. Dictation Speech to text allows to dictate, record, translate and transcribe text instead of typing. It uses latest speech to text voice recognition technology and its main purpose is speech to text and translation for text messaging. Never type any text, just dictate and translate using your speech! Nearly every app that can send text messages can be configured to operate with 'Dictation Speech to text'. Dictate uses the builtin speech to text recognition engine. Dictation Speech to text supports more than 40 languages. Dictate offers 3 text zones, indicated by language flags, for which you can configure a different language in the settings. Thus you can switch between different language projects with a singe click. Translation is as easy as pushing the translation button. You can specify the translation target language in the app settings.
Learn more
Wispr Flow
Wispr Flow is an AI voice-to-text app that turns natural speech into clear, polished writing across apps and devices. The platform lets users dictate messages, documents, code, notes, emails, and replies up to four times faster than typing. Wispr Flow automatically removes filler words, fixes typos, improves formatting, and transforms rambling speech into clean written text. It works across Mac, Windows, iPhone, and Android, helping users write wherever they work. The platform includes AI auto-edits, a personal dictionary, snippet shortcuts, multilingual transcription, and support for more than 100 languages. Built for professionals, creators, students, developers, accessibility users, and teams, Wispr Flow helps people capture ideas faster and write more naturally with their voice.
Learn more
Spokenly
Spokenly is an AI-powered dictation app for Mac, iPhone, Windows, and Linux that turns speech into clean, punctuated text wherever you work. Hold a shortcut, speak naturally, and release to place the transcription directly at the cursor in browsers, email, chat, word processors, IDEs, terminals, and other apps. It supports more than 100 languages, including mixed-language dictation, and offers both local and cloud speech-to-text models. Whisper, Parakeet, and other on-device models can run completely offline, while cloud engines from providers such as OpenAI, Deepgram, Groq, Soniox, and ElevenLabs can be used for higher-accuracy or real-time transcription. Local Only Mode blocks network requests so voice data stays on the device. Modes let users save different transcription models, AI providers, prompts, and output styles for specific tasks, while AI Instructions can remove filler words, fix grammar and punctuation, summarize, rewrite, translate, or reformat dictated text.
Learn more
Amazon Transcribe
Amazon Transcribe makes it easy for developers to add speech to text capabilities to their applications. Audio data is virtually impossible for computers to search and analyze. Therefore, recorded speech needs to be converted to text before it can be used in applications. Historically, customers had to work with transcription providers that required them to sign expensive contracts and were hard to integrate into their technology stacks to accomplish this task. Many of these providers use outdated technology that does not adapt well to different scenarios, like low-fidelity phone audio common in contact centers, which results in poor accuracy. Amazon Transcribe uses a deep learning process called automatic speech recognition (ASR) to convert speech to text quickly and accurately. Amazon Transcribe can be used to transcribe customer service calls, automate subtitling, and generate metadata for media assets to create a fully searchable archive.
Learn more