Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
Interface for OuteTTS models
An Open Source text-to-speech system built by inverting Whisper
A TTS model capable of generating ultra-realistic dialogue
Generate audiobooks from e-books
A generative speech model for daily dialogue
TTS model capable of streaming conversational audio in realtime
A fast TTS architecture with conditional flow matching
A nearly-live implementation of OpenAI's Whisper
SoTA open-source TTS
SOTA Open Source TTS
Capable of understanding text, audio, vision, video
Converts text to speech in realtime
MARS5 speech model (TTS) from CAMB.AI
A Conversational Speech Generation Model
One-click deployment (including offline integration package)
LLM-based Reinforcement Learning audio edit model
Edge TTS Desktop turns text into speech through edge-tts.
Two Integrated Text To Speech Engines uses MMS & Silero
VITS2 backbone with multilingual-bert
Open source implementation of Microsoft's VALL-E X zero-shot TTS model
AI powered speech denoising and enhancement
Best practice TTS based on BERT and VITS
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Implementation of a Transformer based neural network