A simple, high-quality voice conversion tool focused on ease of use
A Web UI for easy subtitle using whisper model
Library for OCR-related tasks powered by Deep Learning
A simple native web interface that uses ChatTTS to synthesize text
Self-host the powerful Chatterbox TTS model
Use Microsoft Edge's online text-to-speech service from Python
A nearly-live implementation of OpenAI's Whisper
Speech-AI-Forge is a project developed around TTS generation model
This does for Documents what repo-browser does for repos
Quick illustration of how one can easily read books together with LLMs
A sound cloning tool with a web interface, using your voice
Stable Diffusion web UI
Context-aware desktop AI assistant that understands screen content
A fast TTS architecture with conditional flow matching
Fast-stable-diffusion + DreamBooth
Automate native Android apps with AI using accessibility APIs
Stable Diffusion WebUI optimized for AMD GPUs with editing tools
Your gateway to GPT writing