Build Vision Agents quickly with any model or video provider
TTS with kokoro and onnx runtime
A simple, high-quality voice conversion tool focused on ease of use
SoTA open-source TTS
Generate audiobooks from e-books, voice cloning & 1107+ languages
SOTA Open Source TTS
Offline Text To Speech synthesis for python
Use Microsoft Edge's online text-to-speech service from Python
High-Quality Voice Cloning TTS for 600+ Languages
A generative speech model for daily dialogue
A nearly-live implementation of OpenAI's Whisper
Comprehensive Gradio WebUI for audio processing
Qwen3-TTS is an open-source series of TTS models
Tokenizer-Free TTS for Multilingual Speech Generation
A simple native web interface that uses ChatTTS to synthesize text
Generate audiobooks from EPUBs, PDFs and text with captions
Instant voice cloning by MIT and MyShell. Audio foundation model
EPUB to audiobook converter, optimized for Audiobookshelf
Industrial-level controllable zero-shot text-to-speech system
A high-quality rapid TTS voice cloning model
The official Python SDK for the ElevenLabs API
Generate audiobooks from e-books
State-of-the-art TTS model under 25MB
A TTS that fits in your CPU (and pocket)
Offline inference engine for art, real-time voice conversations