Image inpainting tool powered by SOTA AI Model
OCR software, free and offline
Faster Whisper transcription with CTranslate2
Generate audiobooks from EPUBs, PDFs and text with captions
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Comprehensive Gradio WebUI for audio processing
Robust Speech Recognition via Large-Scale Weak Supervision
Use Microsoft Edge's online text-to-speech service from Python
A TTS that fits in your CPU (and pocket)
Cut videos with a text editor
Open source healthcare AI
1 min voice data can also be used to train a good TTS model
Stable Diffusion web UI
Unlimited, private and free Speech-To-Text program
A modular voice assistant application for experimenting
Contexts Optical Compression
AsrTools: Smart Voice-to-Text Tool
Translate the video from one language to another and embed dubbing
PDF to Markdown with vision models
EPUB to audiobook converter, optimized for Audiobookshelf
Python library and CLI tool to interface with Google Translate
Visual Causal Flow
OCR model for complex documents with layout-aware structured outputs
A theme for Sublime Text 3 by Mattia Astorino
95% token savings. 155x faster queries. 16 languages