A fast TTS architecture with conditional flow matching
Style-Bert-VITS2: Bert-VITS2 with more controllable voice styles
Controllable and fast Text-to-Speech for over 7000 languages
Foundational model for human-like, expressive TTS
VITS2 backbone with multilingual-bert
A list of accessible speech corpora for ASR, TTS
Implementation of a Transformer based neural network
Deep learning for text to speech
DeepMind's Tacotron-2 Tensorflow implementation
TensorFlow Implementation of DC-TTS: yet another text-to-speech model