Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
A text-to-speech, speech-to-text and speech-to-speech library
High-Quality Voice Cloning TTS for 600+ Languages
Awesome multilingual OCR toolkits based on PaddlePaddle
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Official inference repo for FLUX.1 models
An extremely fast Python type checker and language server
Comprehensive Gradio WebUI for audio processing
Extensions for Python Markdown
A robust, efficient, low-latency speech-to-text library
Generate audiobooks from EPUBs, PDFs and text with captions
Comprehensive Markdown plugin built for Django
State-of-the-art TTS model under 25MB
Python & command-line tool to gather text on the Web
AI bridge enabling assistants to control and automate Unity Editor
A TTS that fits in your CPU (and pocket)
ASCII art library for Python
Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts
A simple native web interface that uses ChatTTS to synthesize text
Automatic Speech Recognition with Word-level Timestamps
EPUB to audiobook converter, optimized for Audiobookshelf
A lightweight approach to removing Google web service dependency
No-code in the front, Python in the back. An open-source framework
Offline inference engine for art, real-time voice conversations
The behavior guidance framework for customer-facing LLM agents