Contexts Optical Compression
OCRmyPDF adds an OCR text layer to scanned PDF files
Awesome multilingual OCR toolkits based on PaddlePaddle
A framework to enable multimodal models to operate a computer
OCR software, free and offline
Visual Causal Flow
Accurate × Fast × Comprehensive
An on-premises, OCR-free unstructured data extraction
A ranked list of awesome machine learning Python libraries
Enhances Tesseract OCR output using LLMs (local or API)
OCR expert VLM powered by Hunyuan's native multimodal architecture
Optical-packet node transceiver frequency allocation
A Python application to add watermarks (text or image) to PDF files
Implementation of Nougat Neural Optical Understanding
Img2Txt - Extract Text From Images using AI
The ultimate tool to automate custom telegram message forwarding
Generate text images for training deep learning ocr model
Constantly summarizing open source dataset and critical papers
Code that accompanies my blog post outlining five video classification