Open speech-to-speech models and pipelines by Hugging Face toolkit AI
Robust Speech Recognition via Large-Scale Weak Supervision
Speech-to-text, text-to-speech, and speaker recognition
Speech recognition module for Python
Multilingual speech recognition and audio understanding model
Open-source industrial-grade ASR models
A PyTorch-based Speech Toolkit
kaldi-asr/kaldi is the official location of the Kaldi project
Audio foundation model excelling in audio understanding
On-device Speech Recognition for Apple Silicon
Captcha solver extension for humans
Speech recognition for your site
Fast and accurate automatic speech recognition (ASR) for edge devices
Port of OpenAI's Whisper model in C/C++
Multilingual Automatic Speech Recognition with word-level timestamps
A free, open source, and extensible speech-to-text application
Automatic Speech Recognition with Word-level Timestamps
Cross-platform AI language practice app
Faster Whisper transcription with CTranslate2
StreamSpeech is a seamless model for offline speech recognition
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
Robust Speech Recognition Across Languages, Dialects
Fast multimodal LLM for real-time voice interaction and AI apps
A cross-platform software for text translation and recognition
Run local LLMs like llama, deepseek, kokoro etc. inside your browser