Framework for building neural networks
Industrial-level controllable zero-shot text-to-speech system
Unofficial Parallel WaveGAN
TTS with kokoro and onnx runtime
A fast TTS architecture with conditional flow matching
SOTA discrete acoustic codec models with 40/75 tokens per second
End-to-end speech processing toolkit
A TTS model capable of generating ultra-realistic dialogue
A webui for different audio related Neural Networks
WaveRNN Vocoder + TTS
Real-Time State-of-the-art Speech Synthesis for Tensorflow 2
Implementation of a Transformer based neural network
Conditional Variational Autoencoder with Adversarial Learning
Generative Adversarial Networks for Efficient and High Fidelity Speech
DeepMind's Tacotron-2 Tensorflow implementation
Toolkit for efficient experimentation with Speech Recognition