Instant voice cloning by MIT and MyShell. Audio foundation model
Towards Human-Sounding Speech
Interface for OuteTTS models
A lightweight text-to-speech model with zero-shot voice cloning
Free, high-quality text-to-speech API endpoint to replace OpenAI
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Self-host the powerful Chatterbox TTS model
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Build Vision Agents quickly with any model or video provider
Tokenizer-Free TTS for Multilingual Speech Generation
Scalable generative AI framework built for researchers and developers
Toolkit for conversational AI
Official MiniMax Model Context Protocol (MCP) server
Converts text to speech in realtime
StreamSpeech is a seamless model for offline speech recognition
An Open Source text-to-speech system built by inverting Whisper
Automatically translates the text of a video based on a subtitle file
End-to-end speech processing toolkit
Multi-lingual large voice generation model, providing inference
Provides CTP stock options and Zhongtai Securities XTP
Bailing is a voice dialogue robot similar to GPT-4o
A TTS model capable of generating ultra-realistic dialogue
Real-time voice interactive digital human
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Lightning-fast, on-device TTS, running natively via ONNX