A simple native web interface that uses ChatTTS to synthesize text
A simple, high-quality voice conversion tool focused on ease of use
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
A high-quality rapid TTS voice cloning model
Industrial-level controllable zero-shot text-to-speech system
High-Quality Voice Cloning TTS for 600+ Languages
Real-time voice interactive digital human
Tokenizer-Free TTS for Multilingual Speech Generation
Offline inference engine for art, real-time voice conversations
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
Spark-TTS Inference Code
Speech-AI-Forge is a project developed around TTS generation model
An Open Source text-to-speech system built by inverting Whisper
StreamSpeech is a seamless model for offline speech recognition
Build Vision Agents quickly with any model or video provider
Converts text to speech in realtime
Scalable generative AI framework built for researchers and developers
Towards Human-Sounding Speech
Bailing is a voice dialogue robot similar to GPT-4o
Controllable & emotion-expressive zero-shot TTS
A sound cloning tool with a web interface, using your voice
One-click deployment (including offline integration package)
Mice speech to text with MX Cinnamon OS ISO
Synchronized Translation for Videos
Towards Human-Level Text-to-Speech through Style Diffusion