Zero-copy PDF text extraction library written in Zig
Regex pattern directory search tool that respects your .gitignore
To extract main article from given URL with Node.js
Fast Rust library for PDF inspection, classification
RAG-Anything: All-in-One RAG Framework
Golang PDF library for creating and processing PDF files (pure go)
A Python library for extracting structured information
Skills, a Chinese software copyright application material generator
A modern PDF library for TypeScript
Build AI-powered semantic search applications
Parser generator to read, process, or translate structured text
Java Pdf Table extraction library
Award-winning modern data processing SDK in C++20
Convert files like docx, xlsx, pptx, html, and more to MarkDown
Blind&Invisible Watermark, image blind watermark, extract watermark
Rust / WASM library for reading, writing and rendering PDF
An extremely fast implementation of Aho Corasick algorithm
Clojure library that parses text into structured data
Expression pattern collection of Chinese compound event extraction
PDF Library for Developers
Summaries and notes on Deep Learning research papers
TextTeaser is an automatic summarization algorithm