State-of-the-art TTS model under 25MB
Qwen3-TTS is an open-source series of TTS models
Instant voice cloning by MIT and MyShell. Audio foundation model
C++ inference library for multiple SVC/TTS
Converts text to speech in realtime
Towards Human-Sounding Speech
A high-quality rapid TTS voice cloning model
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Multi-lingual large voice generation model, providing inference
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Long-form streaming TTS system for multi-speaker dialogue generation
Framework for building neural networks
Controllable and fast Text-to-Speech for over 7000 languages
Easy AI Softwares for Blind, Deaf, Handicapped, Disabled People
Virtual AI anchor that combines state-of-the-art technology
Free & Easy AI Voice Accounting Software For Blind & Speechless People
Toolkit for audio, music, and speech generation
Chinese voice dialogue robot/smart speaker project
General Speech Restoration
Real-Time State-of-the-art Speech Synthesis for Tensorflow 2
Conditional Variational Autoencoder with Adversarial Learning
Generative Adversarial Networks for Efficient and High Fidelity Speech
ColdFusion SDK for the VoiceShot API.
PHP SDK for processing phone calls and SMS through the VoiceShot API.
.NET SDK for processing phone calls and SMS through the VoiceShot API.