Clone a voice in 5 seconds to generate arbitrary speech in real-time
Instant voice cloning by MIT and MyShell. Audio foundation model
A simple, high-quality voice conversion tool focused on ease of use
VoiceStudio is the open-source, fully-local ElevenLabs alternative
Generate audiobooks from e-books, voice cloning & 1107+ languages
A high-quality rapid TTS voice cloning model
Industrial-level controllable zero-shot text-to-speech system
Clone with Python! Data structures for double stranded DNA
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
High-Quality Voice Cloning TTS for 600+ Languages
1 min voice data can also be used to train a good TTS model
MOSS‑TTS Family open‑source speech and sound generation model
A lightweight text-to-speech model with zero-shot voice cloning
Official PyTorch Implementation
Self-host the powerful Chatterbox TTS model
MOSS-TTS-Nano is an open-source multilingual tiny speech generation
NeuTTS model built from small LLM backbones
On-device TTS model by Neuphonic
Tokenizer-Free TTS for Multilingual Speech Generation
Open-source framework for intelligent speech interaction
Browser userscript that enhances ChatGPT reliability and usability
Community-maintained approach to improving access to GitHub services
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
The official Python SDK for the ElevenLabs API
Spark-TTS Inference Code