Zero-copy PDF text extraction library written in Zig
To extract main article from given URL with Node.js
Fast Rust library for PDF inspection, classification
RAG-Anything: All-in-One RAG Framework
A Python library for extracting structured information
A modern PDF library for TypeScript
Blind&Invisible Watermark, image blind watermark, extract watermark
Rust / WASM library for reading, writing and rendering PDF
An extremely fast implementation of Aho Corasick algorithm
Clojure library that parses text into structured data
Expression pattern collection of Chinese compound event extraction
PDF Library for Developers
Summaries and notes on Deep Learning research papers
TextTeaser is an automatic summarization algorithm
Smarter YAML front matter parser