OCR software, free and offline
Contexts Optical Compression
PDF to Markdown with vision models
Accurate × Fast × Comprehensive
Welcome the Era of One-shot Long-horizon Parsing
Visual Causal Flow
OCRmyPDF adds an OCR text layer to scanned PDF files
Formula recognition based on LaTeX-OCR and ONNXRuntime
Awesome multilingual OCR toolkits based on PaddlePaddle
An Open-Source Toolkit for General-OCR Research and Applications
Enhances Tesseract OCR output using LLMs (local or API)
A high-quality tool for convert PDF to Markdown and JSON
OCR expert VLM powered by Hunyuan's native multimodal architecture
Ready-to-use OCR with 80+ supported languages
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Library for OCR-related tasks powered by Deep Learning
PDF scientific paper translation with preserved formats
Windrecorder is a memory search app by records everything
Multilingual Document Layout Parsing in a Single Vision-Language Model
Math OCR model that outputs LaTeX and markdown
A framework to enable multimodal models to operate a computer
Open Source Document Management System for Digital Archives
Make any agent harness multimodal-native
Convert AI papers to GUI
A simple tool for reading in poorly redacted documents