A lightweight audio-to-MIDI converter with pitch bend detection
Industrial-level controllable zero-shot text-to-speech system
Qwen3-TTS is an open-source series of TTS models
Use Microsoft Edge's online text-to-speech service from Python
Spark-TTS Inference Code
Open source text-to-speech tool, supports extra-long text
Audio Plugin for Audio to MIDI transcription using deep learning
High-Quality Voice Cloning TTS for 600+ Languages
Claude code for everything except coding
C++ inference library for multiple SVC/TTS
Transform your voice in real-time voxal voice changer
An AI for Music Generation
Towards Human-Level Text-to-Speech through Style Diffusion
Multi-Voice and Prompt-Controlled TTS Engine
Open source implementation of Microsoft's VALL-E X zero-shot TTS model
Simple and powerful voice changer for Linux, written with Python & GTK
SoftVC VITS Singing Voice Conversion
Singing Voice Synthesis via Shallow Diffusion Mechanism
Task of transcribing piano recordings into MIDI files
An awesome browser extension that reads aloud webpage content
List of useful data augmentation resources
This project includes basic NLP and DSP techniques for Text-to-Speech