Minimalistic audiobook player
Comprehensive Gradio WebUI for audio processing
Fast and accurate automatic speech recognition (ASR) for edge devices
MEGA Android App
Speech to Text to Speech, sends text as OSC messages
GLM-4-Voice | End-to-End Chinese-English Conversational Model
A sound cloning tool with a web interface, using your voice
A fan port of Cave Story for the Sega Mega Drive
The open-source voice synthesis studio powered by Qwen3-TTS
Documents and exposes generated source for menu-bar workflows
Instant voice cloning by MIT and MyShell. Audio foundation model
Clone a voice in 5 seconds to generate arbitrary speech in real-time
A simple, high-quality voice conversion tool focused on ease of use
High-Quality Voice Cloning TTS for 600+ Languages
Industrial-level controllable zero-shot text-to-speech system
Telegram Desktop messaging app
In-App assistant SDK to build a multimodal conversational UX websites
1 min voice data can also be used to train a good TTS model
Conversational voice AI agents
AI-polished text appears at your cursor in any app
Framework for building real-time voice and multimodal AI agents
OpenVoiceOS Core, the FOSS Artificial Intelligence platform
Official PyTorch Implementation
On-device wake word detection powered by deep learning
Qwen3-TTS is an open-source series of TTS models