A Python tool to help extracting information from structured PDFs
Simple LaTeX parser providing latex-to-unicode and unicode-to-latex
Bibtex parser for Python 3
A fast, extensible and spec-compliant Markdown parser in pure Python
Open-Source Python3 tool for recognizing layouts, tables, and math
OCRmyPDF adds an OCR text layer to scanned PDF files
CLI tool and python library
Edit PDF files with Nano Banana
A Python utility / library to sort imports
A simple tool for reading in poorly redacted documents
A fast and minimal circular JSON parser
Video-based AI memory library. Store millions of text chunks in MP4
The next generation Javascript WYSIWYG HTML Editor
Math OCR model that outputs LaTeX and markdown
Tools to ease the creation of snippets, syntax definitions, etc.
JSON Lint for PHP
CLI tool to extract (meta)data from PDF and manipulate PDF files
Re-editable LaTeX/ typst graphics for Inkscape
A toolchain for web projects, aimed to provide functionalities
Extract one time password (OTP) secrets from QR codes
PDF Parser for AI-ready data. Automate PDF accessibility
Java library for working with real-world HTML
The data structure for multimodal data
HTML Loader
Converts CSS selectors to XPath expressions