1 min voice data can also be used to train a good TTS model
SoTA open-source TTS
A generative speech model for daily dialogue
Open-source framework for intelligent speech interaction
Towards Human-Sounding Speech
FAIR Sequence Modeling Toolkit 2
An Open Source text-to-speech system built by inverting Whisper
Python library and CLI tool to interface with Google Translate
Foundational model for human-like, expressive TTS
VITS2 backbone with multilingual-bert
AI powered speech denoising and enhancement
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Implementation of a Transformer based neural network
TensorFlow Implementation of DC-TTS: yet another text-to-speech model