State-of-the-art TTS model under 25MB
Qwen3-TTS is an open-source series of TTS models
Instant voice cloning by MIT and MyShell. Audio foundation model
C++ inference library for multiple SVC/TTS
Converts text to speech in realtime
Towards Human-Sounding Speech
A high-quality rapid TTS voice cloning model
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Multi-lingual large voice generation model, providing inference
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Long-form streaming TTS system for multi-speaker dialogue generation
Framework for building neural networks
Controllable and fast Text-to-Speech for over 7000 languages
Mice speech to text with MX Cinnamon OS ISO
Easy AI Softwares for Blind, Deaf, Handicapped, Disabled People
Virtual AI anchor that combines state-of-the-art technology
Free & Easy AI Voice Accounting Software For Blind & Speechless People
Toolkit for audio, music, and speech generation
A fast, local neural text to speech system
Chinese voice dialogue robot/smart speaker project
General Speech Restoration
Real-Time State-of-the-art Speech Synthesis for Tensorflow 2
Conditional Variational Autoencoder with Adversarial Learning
Generative Adversarial Networks for Efficient and High Fidelity Speech