OCR software, free and offline
Accurate × Fast × Comprehensive
Contexts Optical Compression
Visual Causal Flow
Welcome the Era of One-shot Long-horizon Parsing
OCRmyPDF adds an OCR text layer to scanned PDF files
Awesome multilingual OCR toolkits based on PaddlePaddle
An Open-Source Toolkit for General-OCR Research and Applications
Enhances Tesseract OCR output using LLMs (local or API)
A high-quality tool for convert PDF to Markdown and JSON
OCR expert VLM powered by Hunyuan's native multimodal architecture
Multilingual Document Layout Parsing in a Single Vision-Language Model
Convert AI papers to GUI
Get your documents ready for gen AI
A framework to enable multimodal models to operate a computer
Open Source Document Management System for Digital Archives
OpenRecall is a fully open-source, privacy-first alternative
A Repo For Document AI
Document content and metadata extraction microservice
Structured data extraction and instruction calling with ML, LLM
OCR model for complex documents with layout-aware structured outputs
A community-supported supercharged version of paperless
AI tool for automating desktop tasks via natural language input
An on-premises, OCR-free unstructured data extraction
In-depth tutorials on LLMs, RAGs and real-world AI agent applications