RiverScript
Transcribe everything you can hear on your computer
Capture and turn into text everything you can hear on your computer – meetings, podcasts, any videos with Live Recording Transcription from RiverScript. Your sound – your rules. A multi-model AI architecture combining leading speech recognition models from ElevenLabs, OpenAI and Deepgram. Interactive editor, timecodes, speaker diarization. Lightning-fast desktop client for Windows and macOS, built on Rust. Supports audio and video files up to 50 GB and 8 hours long.
● works with audio and video files up to 50 GB, including batch uploads
● has a built-in editor and an interactive media player
● translates transcripts into other languages with AI
● generates subtitles with clickable timestamps
● performs speaker diarization
● creates AI-powered summaries
● lets you ask AI anything about your transcript
RiverScript – transcribe everything!
Learn more
Temi
Upload any audio or video file. We accept all file types. Review your transcript with timestamps and speakers. Save & export your transcript as MS Word, PDF, SRT, VTT and more. Transcript quality depends on audio quality. Record clear audio to get accurate transcripts. Temi's free transcription editor lets you edit your transcripts online in minutes. Built by our machine learning and speech recognition experts. Quickly clean-up the provided transcript. Adjust the playback speed and skip around easily. Temi knows the timing of every word. Add any timestamps. We mark the change of every speaker and label them. Download your transcript into text (MS Word, PDF) or closed caption files (SRT, VTT).
Learn more
Silkwave Voice
Silkwave Voice is a privacy-focused audio recording and transcription app for macOS. Record from your microphone, system audio, or both at once - with accurate, real-time transcription powered by Apple's on-device speech-to-text models. No cloud uploads, no subscriptions, no per-minute API costs.
RECORD ANY AUDIO SOURCE
• Microphone - voice notes, in-person meetings, dictation
• System Audio - Zoom, Google Meet, Teams, YouTube, browser tabs
• Both at once - capture your mic and remote participants simultaneously
ON-DEVICE TRANSCRIPTION
• Real-time speech-to-text using Apple's on-device models
• 10 languages: Cantonese, Chinese, English, French, German, Italian, Japanese, Korean, Portuguese, Spanish
• Completely local - no internet connection needed
AI-POWERED SUMMARIES
• Structured summaries with key topics, action items, and decisions
• Powered by ChatGPT through Apple Intelligence - no API keys needed
Learn more
Subanana
Subanana is an AI speech-to-text web app that turns audio and video into subtitles, transcripts, and meeting summaries in 80+ languages, with standout accuracy on Asian and mixed-language speech (Cantonese, Mandarin, Japanese, Korean, and code-switching) that English-first tools handle poorly.
Subtitles: import a file or a YouTube/Instagram/Facebook link, edit with a glossary and AI auto-correct, and export SRT, VTT, TXT, DOCX, bilingual subtitles, or burned-in video.
Transcripts: speaker labels, filler-word removal, automatic punctuation and paragraphs.
Meeting summaries: templates, decisions and action items, plus a Google Meet and Microsoft Teams recording bot that processes the meeting after it ends.
Live captions: real-time captioning with translation for events.
Learn more