Document (PDF, Word, PPTX ...) extraction and parse API
Extract one time password (OTP) secrets from QR codes
Read and extract text and other content from PDFs in C#
A pure-python PDF library capable of splitting, merging, cropping
WindowTextExtractor allows you to get a text from any OS
Library for OCR-related tasks powered by Deep Learning
JavaScript OCR and text extraction for images and PDFs
PDFsam, a desktop application to split, merge, mix, rotate PDF files
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
Comprehensive Gradio WebUI for audio processing
A cross-platform software for text translation and recognition
OCR software, free and offline
Image Toolbox is an powerful picture editor, which can crop
OCR model for complex documents with layout-aware structured outputs
Contexts Optical Compression
A simple native web interface that uses ChatTTS to synthesize text
Ksoup is a lightweight Kotlin Multiplatform library for parsing HTML
A fast, helpful, and open-source document parser
Handwritten Text Recognition (HTR) system implemented with TensorFlow
The Refactoring library based off the Refactoring book
Extract structured data from webpages using LLM-powered scraping
Python bindings for MuPDF's rendering library.
Open source semantic search and text analytics for large document sets
Generate blog articles from video or audio
A modular graph-based Retrieval-Augmented Generation (RAG) system