Edit PDF files with Nano Banana
OCRmyPDF adds an OCR text layer to scanned PDF files
Document (PDF, Word, PPTX ...) extraction and parse API
Python bindings for MuPDF's rendering library.
A pure-python PDF library capable of splitting, merging, cropping
CLI tool to extract (meta)data from PDF and manipulate PDF files
borb is a library for reading, creating and manipulating PDF files
Zero-copy PDF text extraction library written in Zig
OCR software, free and offline
A high-quality PDF to Markdown tool based on large language model
A library for converting HTML into PDFs using ReportLab
A community-supported supercharged version of paperless
Ferramenta de Tarjamento de Dados Pessoais e Sigilosos
Open-Source Python3 tool for recognizing layouts, tables, and math
A simple tool for reading in poorly redacted documents
Reading book source
Generate audiobooks from EPUBs, PDFs and text with captions
The best free open source website change detection and restock service
Googles NotebookLM but local
Make bilingual epub books Using AI translate
High accuracy RAG for answering questions from scientific documents
PDF to Markdown with vision models
Run Bonsai (1-bit) and Ternary-Bonsai language models locally
Open Source Document Management System for Digital Archives
The Markdown Editor for Linux