NeuTTS model built from small LLM backbones
Controllable & emotion-expressive zero-shot TTS
On-device TTS model by Neuphonic
GLM-4-Voice | End-to-End Chinese-English Conversational Model
Converts text to speech in realtime
Open-source multi-speaker long-form text-to-speech model
Capable of understanding text, audio, vision, video
Towards Human-Sounding Speech
Speech-AI-Forge is a project developed around TTS generation model
LLM-based Reinforcement Learning audio edit model
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)