An Open Source text-to-speech system built by inverting Whisper
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
NeuTTS model built from small LLM backbones
On-device TTS model by Neuphonic
TTS model capable of streaming conversational audio in realtime
The official Python library for the Fish Audio API
MOSS‑TTS Family open‑source speech and sound generation model
Long-form streaming TTS system for multi-speaker dialogue generation
A TTS model capable of generating ultra-realistic dialogue
Open-source framework for intelligent speech interaction
Speech-AI-Forge is a project developed around TTS generation model
Converts text to speech in realtime
Miso TTS is an 8 billion, highly emotive text-to-speech model
LLM-based Reinforcement Learning audio edit model
Capable of understanding text, audio, vision, video
A sound cloning tool with a web interface, using your voice
A fast TTS architecture with conditional flow matching
Foundational model for human-like, expressive TTS
MARS5 speech model (TTS) from CAMB.AI
One-click deployment (including offline integration package)
A Conversational Speech Generation Model
Two Integrated Text To Speech Engines uses MMS & Silero
Towards Human-Level Text-to-Speech through Style Diffusion
VITS2 backbone with multilingual-bert
AI powered speech denoising and enhancement