Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Chuyển đổi văn bản thành giọng nói không giới hạn
Miso TTS is an 8 billion, highly emotive text-to-speech model
Self-host the powerful Chatterbox TTS model
Offline Text To Speech synthesis for python
Use Microsoft Edge's online text-to-speech service from Python
Long-form streaming TTS system for multi-speaker dialogue generation
Speech recognition module for Python
A high-quality rapid TTS voice cloning model
Faster Whisper transcription with CTranslate2
Free, high-quality text-to-speech API endpoint to replace OpenAI
MOSS‑TTS Family open‑source speech and sound generation model
The official Python SDK for the ElevenLabs API
A nearly-live implementation of OpenAI's Whisper
MOSS-TTS-Nano is an open-source multilingual tiny speech generation
An Open Source text-to-speech system built by inverting Whisper
Capable of understanding text, audio, vision, video
A TTS model capable of generating ultra-realistic dialogue
NeuTTS model built from small LLM backbones
Spark-TTS Inference Code
SoTA open-source TTS
Open-source multi-speaker long-form text-to-speech model
Multi-lingual large voice generation model, providing inference
Open-source framework for intelligent speech interaction
Qwen3-omni is a natively end-to-end, omni-modal LLM