The official Python library for the Fish Audio API
TTS with kokoro and onnx runtime
State-of-the-art TTS model under 25MB
1 min voice data can also be used to train a good TTS model
A generative speech model for daily dialogue
Qwen3-TTS is an open-source series of TTS models
High-Quality Voice Cloning TTS for 600+ Languages
MOSS-TTS-Nano is an open-source multilingual tiny speech generation
Instant voice cloning by MIT and MyShell. Audio foundation model
A nearly-live implementation of OpenAI's Whisper
A TTS that fits in your CPU (and pocket)
Industrial-level controllable zero-shot text-to-speech system
SoTA open-source TTS
Long-form streaming TTS system for multi-speaker dialogue generation
An inference library for Kokoro-82M
A lightweight text-to-speech model with zero-shot voice cloning
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Tokenizer-Free TTS for Multilingual Speech Generation
Generate audiobooks from e-books
A sound cloning tool with a web interface, using your voice
NeuTTS model built from small LLM backbones
Python library and CLI tool to interface with Google Translate
SOTA Open Source TTS
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Multi-lingual large voice generation model, providing inference