Controllable & emotion-expressive zero-shot TTS
An Open Source text-to-speech system built by inverting Whisper
MARS5 speech model (TTS) from CAMB.AI
Towards Human-Level Text-to-Speech through Style Diffusion
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Conditional Variational Autoencoder with Adversarial Learning
TensorFlow Implementation of DC-TTS: yet another text-to-speech model