A cross-platform software for text translation and recognition
Spark-TTS Inference Code
Production ready toolkit to run AI locally
Lightning-fast, on-device TTS, running natively via ONNX
Free, high-quality text-to-speech API endpoint to replace OpenAI
ComfyUI integration for Microsoft's VibeVoice text-to-speech model
Speakr is a personal, self-hosted web application
A high-quality rapid TTS voice cloning model
Open-source framework for intelligent speech interaction
PersonaPlex code
Generate audiobooks from e-books
AI-polished text appears at your cursor in any app
Like the macOS say command, but with a modern voice
Qwen3-omni is a natively end-to-end, omni-modal LLM
The behavior guidance framework for customer-facing LLM agents
Open source text-to-speech tool, supports extra-long text
A nearly-live implementation of OpenAI's Whisper
TTS model capable of streaming conversational audio in realtime
Free open source speech synthesizer for Russian and other languages
NeuTTS model built from small LLM backbones
On-device TTS model by Neuphonic
End-to-end speech processing toolkit
Instant voice cloning by MIT and MyShell. Audio foundation model
TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning
Voice Recognition to Text Tool