Controllable & emotion-expressive zero-shot TTS
LLM-based Reinforcement Learning audio edit model
Open-source multi-speaker long-form text-to-speech model
FAIR Sequence Modeling Toolkit 2
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
A python package to analyze and compare voices with deep learning
TensorFlow Implementation of DC-TTS: yet another text-to-speech model