Multi-lingual large voice generation model, providing inference
An Open Source text-to-speech system built by inverting Whisper
SOTA Open Source TTS
Instant voice cloning by MIT and MyShell. Audio foundation model
TTS model capable of streaming conversational audio in realtime
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
FAIR Sequence Modeling Toolkit 2
Lightning-fast, on-device TTS, running natively via ONNX
MOSS‑TTS Family open‑source speech and sound generation model
Long-form streaming TTS system for multi-speaker dialogue generation
A TTS model capable of generating ultra-realistic dialogue
LLM-based Reinforcement Learning audio edit model
Two Integrated Text To Speech Engines uses MMS & Silero
A fast, local neural text to speech system
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Conditional Variational Autoencoder with Adversarial Learning
Deep learning for text to speech