Open speech-to-speech models and pipelines by Hugging Face toolkit AI
Speech Note Linux app. Note taking, reading and translating
Robust Speech Recognition via Large-Scale Weak Supervision
Doing Phonetics By Computer
End-to-end speech processing toolkit
Translate the video from one language to another and embed dubbing
Fast and accurate automatic speech recognition (ASR) for edge devices
VoiceStudio is the open-source, fully-local ElevenLabs alternative
Toolkit for conversational AI
A free, open source, and extensible speech-to-text application
Automatic Speech Recognition with Word-level Timestamps
Han Language Processing
Stanford CoreNLP, a Java suite of core NLP tools
Faster Whisper transcription with CTranslate2
Generate audiobooks from EPUBs, PDFs and text with captions
Underthesea - Vietnamese NLP Toolkit
Fastest macOS Offline Dictation app
OpenVINO™ Toolkit repository
Comprehensive Gradio WebUI for audio processing
Persian NLP Toolkit
A modular voice assistant application for experimenting
Fast multimodal LLM for real-time voice interaction and AI apps
Open Source Speech Language Model
Use Microsoft Edge's online text-to-speech service from Python
The best open-source alternative to Superwhisper & Wispr Flow