OCR software, free and offline
Contexts Optical Compression
PDF to Markdown with vision models
Accurate × Fast × Comprehensive
Welcome the Era of One-shot Long-horizon Parsing
Visual Causal Flow
Formula recognition based on LaTeX-OCR and ONNXRuntime
OCRmyPDF adds an OCR text layer to scanned PDF files
An Open-Source Toolkit for General-OCR Research and Applications
Awesome multilingual OCR toolkits based on PaddlePaddle
Enhances Tesseract OCR output using LLMs (local or API)
A high-quality tool for convert PDF to Markdown and JSON
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Library for OCR-related tasks powered by Deep Learning
OCR expert VLM powered by Hunyuan's native multimodal architecture
Multilingual Document Layout Parsing in a Single Vision-Language Model
Windrecorder is a memory search app by records everything
Ready-to-use OCR with 80+ supported languages
PDF scientific paper translation with preserved formats
Math OCR model that outputs LaTeX and markdown
Get your documents ready for gen AI
Open Source Document Management System for Digital Archives
Make any agent harness multimodal-native
A framework to enable multimodal models to operate a computer
A simple tool for reading in poorly redacted documents