Create synth presets from words
Open source implementation of Microsoft's VALL-E X zero-shot TTS model
Search recursively all files, text inside files, and bookmarks
Best practice TTS based on BERT and VITS
Unofficial Parallel WaveGAN
A webui for different audio related Neural Networks
SoftVC VITS Singing Voice Conversion
Implementation of MusicLM music generation model in Pytorch
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
Automatically generate and overlay subtitles for any video
Audio generation using diffusion models, in PyTorch
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Txt-2-Mp3 6.3 Mark 2 [Improved.Simplified.Alternative]
No-code tool for creating a neural search solution in minutes
A walk along memory lane
Implementation of NÜWA, attention network for text to video synthesis
Real-time music generation using stable diffusion techniques AI
Using OpenAI's Whisper to automatically generate YouTube subtitles
Dump psg/ym chip tune files to txt and midi format
WaveRNN Vocoder + TTS
Based on the Disco Diffusion, version of the AI art creation software
State of the art faster Transformer with Tensorflow 2.0
A data augmentations library for audio, image, text, and video
❤️ 18k-youtube-download with python and kivy Dev.Wk-18k
General Speech Restoration