A fast TTS architecture with conditional flow matching
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
Controllable and fast Text-to-Speech for over 7000 languages
Foundational model for human-like, expressive TTS
VITS2 backbone with multilingual-bert
Implementation of a Transformer based neural network
DeepMind's Tacotron-2 Tensorflow implementation
TensorFlow Implementation of DC-TTS: yet another text-to-speech model