Refractoring ChatBot+LLM, Gpt-3.5-turbo, ChatGPT Bot/Voice Assistant
Chat & pretrained large audio language model proposed by Alibaba Cloud
Software that uses AI to perform real-time voice conversion
Toolkit for audio, music, and speech generation
Text to Speech Utility
VITS2 backbone with multilingual-bert
AI powered speech denoising and enhancement
Multi-Voice and Prompt-Controlled TTS Engine
AIlice is a fully autonomous, general-purpose AI agent
A deep learning toolkit for Text-to-Speech, battle-tested in research
Best practice TTS based on BERT and VITS
Open source implementation of Microsoft's VALL-E X zero-shot TTS model
Unofficial Parallel WaveGAN
SoftVC VITS Singing Voice Conversion
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
Chinese voice dialogue robot/smart speaker project
A webui for different audio related Neural Networks
Video automatic transcribe and translated subtitle generator
A GUI tool for generating subtitle from videos, generating srt files
Automatically generate and overlay subtitles for any video
PyTorch implementation of VALL-E (Zero-Shot Text-To-Speech)
Txt-2-Mp3 6.3 Mark 2 [Improved.Simplified.Alternative]
Contextually-keyed word vectors
NLP, before and after spaCy