Chat & pretrained large audio language model proposed by Alibaba Cloud
Text to Speech Utility
Toolkit for audio, music, and speech generation
VITS2 backbone with multilingual-bert
AI powered speech denoising and enhancement
Multi-Voice and Prompt-Controlled TTS Engine
AIlice is a fully autonomous, general-purpose AI agent
Open source implementation of Microsoft's VALL-E X zero-shot TTS model
A deep learning toolkit for Text-to-Speech, battle-tested in research
Best practice TTS based on BERT and VITS
Unofficial Parallel WaveGAN
SoftVC VITS Singing Voice Conversion
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
Chinese voice dialogue robot/smart speaker project
A webui for different audio related Neural Networks
Video automatic transcribe and translated subtitle generator
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Txt-2-Mp3 6.3 Mark 2 [Improved.Simplified.Alternative]
Contextually-keyed word vectors
NLP, before and after spaCy
Frontend and Backend Code for ArtikelSchreiber.com and UNAIQUE.NET
A walk along memory lane
Singing Voice Synthesis via Shallow Diffusion Mechanism
Web mining module for Python, with tools for scraping