TTS with kokoro and onnx runtime
Offline Text To Speech synthesis for python
End-to-end speech processing toolkit
Generate audiobooks from EPUBs, PDFs and text with captions
A robust, efficient, low-latency speech-to-text library
Comprehensive Gradio WebUI for audio processing
Converts text to speech in realtime
Offline inference engine for art, real-time voice conversations
Qwen3-ASR is an open-source series of ASR models
Open source text-to-speech tool, supports extra-long text
Open-Source CapCut replacement (MCP supported)
Hardware-accelerated video transcoding using Android MediaCodec APIs
Statistical machine intelligence and learning engine
On-device Speech-to-Intent engine powered by deep learning
Go efficient multilingual NLP and text segmentation
Java library designed to integrate Speech-to-Text
PDF to Podcast transforms any PDF document into a podcast-ready audio
Local-first AI audio processing, transcription and mastering
Latest Debian and Ubuntu based REPO files and ISO's for PearlLinuxOS
Convert VoIP calls to text and analyze them with AI
Virtual AI anchor that combines state-of-the-art technology
Multi-Voice and Prompt-Controlled TTS Engine
Amica is an open source interface for interactive communication
Chinese text-to-speech engine
AI macOS app for real-time coding interview coaching assistance