TTS with kokoro and onnx runtime
1 min voice data can also be used to train a good TTS model
Industrial-level controllable zero-shot text-to-speech system
High-Quality Voice Cloning TTS for 600+ Languages
State-of-the-art TTS model under 25MB
Generate audiobooks from e-books
Tokenizer-Free TTS for Multilingual Speech Generation
Open-source multi-speaker long-form text-to-speech model
Towards Human-Sounding Speech
Controllable & emotion-expressive zero-shot TTS
Speech-AI-Forge is a project developed around TTS generation model
NeuTTS model built from small LLM backbones
On-device TTS model by Neuphonic
MOSS‑TTS Family open‑source speech and sound generation model
The official Python library for the Fish Audio API
Converts text to speech in realtime
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Miso TTS is an 8 billion, highly emotive text-to-speech model
LLM-based Reinforcement Learning audio edit model
An Open Source text-to-speech system built by inverting Whisper
A fast TTS architecture with conditional flow matching
A sound cloning tool with a web interface, using your voice
One-click deployment (including offline integration package)
Towards Human-Level Text-to-Speech through Style Diffusion
VITS2 backbone with multilingual-bert