OCR software, free and offline
PDF to Markdown with vision models
Accurate × Fast × Comprehensive
Contexts Optical Compression
Visual Causal Flow
Welcome the Era of One-shot Long-horizon Parsing
OCRmyPDF adds an OCR text layer to scanned PDF files
Formula recognition based on LaTeX-OCR and ONNXRuntime
Awesome multilingual OCR toolkits based on PaddlePaddle
An Open-Source Toolkit for General-OCR Research and Applications
A high-quality tool for convert PDF to Markdown and JSON
Enhances Tesseract OCR output using LLMs (local or API)
OCR expert VLM powered by Hunyuan's native multimodal architecture
Multilingual Document Layout Parsing in a Single Vision-Language Model
PDF scientific paper translation with preserved formats
Windrecorder is a memory search app by records everything
Convert AI papers to GUI
Math OCR model that outputs LaTeX and markdown
Get your documents ready for gen AI
A framework to enable multimodal models to operate a computer
Open Source Document Management System for Digital Archives
A simple tool for reading in poorly redacted documents
OpenRecall is a fully open-source, privacy-first alternative
A Repo For Document AI
Document content and metadata extraction microservice