Python library and CLI tool to interface with Google Translate
super expressive prompting model based on ltx2.3
Open Source Speech Language Model
Framework for building realtime multimodal voice AI agents apps
StreamSpeech is a seamless model for offline speech recognition
Cross-platform AI language practice app
Toolkit for conversational AI
Offline Text To Speech synthesis for python
MOSS-TTS-Nano is an open-source multilingual tiny speech generation
Long-form streaming TTS system for multi-speaker dialogue generation
Capable of understanding text, audio, vision, video
GLM-4-Voice | End-to-End Chinese-English Conversational Model
The open-source voice synthesis studio powered by Qwen3-TTS
A robust, efficient, low-latency speech-to-text library
Use Microsoft Edge's online text-to-speech service from Python
Controllable & emotion-expressive zero-shot TTS
Towards Human-Sounding Speech
Self-host the powerful Chatterbox TTS model
Open-source multi-speaker long-form text-to-speech model
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Faster Whisper transcription with CTranslate2
Converts text to speech in realtime
MOSS‑TTS Family open‑source speech and sound generation model
Translate the video from one language to another and embed dubbing
A cross-platform software for text translation and recognition