Offline Text To Speech synthesis for python
Comprehensive Gradio WebUI for audio processing
Audio foundation model excelling in audio understanding
State-of-the-art TTS model under 25MB
Use Microsoft Edge's online text-to-speech service from Python
Open Source Speech Language Model
Open-source framework for intelligent speech interaction
GLM-4-Voice | End-to-End Chinese-English Conversational Model
A PyTorch-based Speech Toolkit
A simple, high-quality voice conversion tool focused on ease of use
Miso TTS is an 8 billion, highly emotive text-to-speech model
A lightweight text-to-speech model with zero-shot voice cloning
Open-source multi-speaker long-form text-to-speech model
1 min voice data can also be used to train a good TTS model
Framework for building realtime multimodal voice AI agents apps
Converts text to speech in realtime
Generate audiobooks from EPUBs, PDFs and text with captions
A high-quality rapid TTS voice cloning model
Controllable & emotion-expressive zero-shot TTS
PersonaPlex code
The official Python SDK for the ElevenLabs API
Python library and CLI tool to interface with Google Translate
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
Powerful Android AI agent with tools, automation, and Linux shell
Towards Human-Sounding Speech