A simple native web interface that uses ChatTTS to synthesize text
Workflow and speech recognition app
A simple, high-quality voice conversion tool focused on ease of use
Code for openai.fm, a demo for the OpenAI Speech API
Cross-platform AI language practice app
A high-quality rapid TTS voice cloning model
A cross-platform software for text translation and recognition
Real-time voice interactive digital human
Industrial-level controllable zero-shot text-to-speech system
High-Quality Voice Cloning TTS for 600+ Languages
Open source text-to-speech tool, supports extra-long text
The python library for real-time communication
Offline inference engine for art, real-time voice conversations
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Tokenizer-Free TTS for Multilingual Speech Generation
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
Spark-TTS Inference Code
Speech-AI-Forge is a project developed around TTS generation model
An Open Source text-to-speech system built by inverting Whisper
StreamSpeech is a seamless model for offline speech recognition
Build Vision Agents quickly with any model or video provider
Converts text to speech in realtime
Scalable generative AI framework built for researchers and developers
Towards Human-Sounding Speech
A single Gradio + React WebUI with extensions for ACE-Step