A simple, high-quality voice conversion tool focused on ease of use
Instant voice cloning by MIT and MyShell. Audio foundation model
Multi-lingual large voice generation model, providing inference
Official PyTorch Implementation
On-device Speech-to-Intent engine powered by deep learning
Spark-TTS Inference Code
MOSS‑TTS Family open‑source speech and sound generation model
PersonaPlex code
Repo of Qwen2-Audio chat & pretrained large audio language model
Removes 20+ patterns of AI slop from any piece of writing
TTS model capable of streaming conversational audio in realtime
Aider is AI pair programming in your terminal
Offline Text To Speech synthesis for python
Open-source Video Translation Skill
A natural language interface for computers
Agent skill: make LLMs write docs in ASD-STE100
A specialized Claude Code workspace for creating long-form
Generate high-definition story short videos with one click using AI
Fully Local Manus AI. No APIs, No $200 monthly bills
Offline inference engine for art, real-time voice conversations
SOTA Open Source TTS
Dockerized FastAPI wrapper for Kokoro-82M text-to-speech model
An Open Source text-to-speech system built by inverting Whisper
Long-form streaming TTS system for multi-speaker dialogue generation
A TTS model capable of generating ultra-realistic dialogue