AI tool that removes hardcoded subtitles and text from videos locally
Open-Source Python3 tool for recognizing layouts, tables, and math
Offline Text To Speech synthesis for python
OCRmyPDF adds an OCR text layer to scanned PDF files
Awesome multilingual OCR toolkits based on PaddlePaddle
Build and connect intelligent bots that interact naturally
Use Microsoft Edge's online text-to-speech service from Python
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
High-Quality Voice Cloning TTS for 600+ Languages
Comprehensive Gradio WebUI for audio processing
Vim Win32 Installer
Official inference repo for FLUX.1 models
State-of-the-art TTS model under 25MB
A generative speech model for daily dialogue
Focus on prompting and generating
FastAPI framework, high performance, easy to learn, fast to code
Contexts Optical Compression
Sphinx source parser for Jupyter notebooks
Automatic Speech Recognition with Word-level Timestamps
Cut videos with a text editor
Nexa SDK is a comprehensive toolkit for supporting ONNX and GGML
A TTS that fits in your CPU (and pocket)
Extensions for Python Markdown
A nearly-live implementation of OpenAI's Whisper
ASCII art library for Python