A walk along memory lane
State-of-the-art deep learning based audio codec
Singing Voice Synthesis via Shallow Diffusion Mechanism
Using OpenAI's Whisper to automatically generate YouTube subtitles
Web mining module for Python, with tools for scraping
A Deep-Learning-Based Chinese Speech Recognition System
WaveRNN Vocoder + TTS
[WIP] VoiceSmith makes training text to speech models easy
State of the art faster Transformer with Tensorflow 2.0
A data augmentations library for audio, image, text, and video
We provide a PyTorch implementation of the paper Voice Separation
A Python/Pytorch app for easily synthesising human voices
A CLI script to generate subtitle files (SRT/VTT/TXT) for any video
Python package for Korean natural language processing
Separate audio recordings into individual sources
Mycroft Core, the Mycroft Artificial Intelligence platform
Main repository of Project Alice, contains main unit source code
Clone a voice in 5 seconds to generate arbitrary speech in real-time
General Speech Restoration
PAddle PARAllel text-to-speech toolKIT
Python AI assistant
Real-Time State-of-the-art Speech Synthesis for Tensorflow 2
Conditional Variational Autoencoder with Adversarial Learning
Kashgari is a production-level NLP Transfer learning framework
Implementation of a Transformer based neural network