Document (PDF, Word, PPTX ...) extraction and parse API
Recognition and resolution of numbers, units, date/time, etc.
JavaScript OCR and text extraction for images and PDFs
Zero-copy PDF text extraction library written in Zig
Python & command-line tool to gather text on the Web
CLI tool to extract (meta)data from PDF and manipulate PDF files
MD/.JSON Document OCR and structured data extraction API
Official Vectorize MCP Server
Fast and efficient unstructured data extraction
Structured data extraction and instruction calling with ML, LLM
A machine learning software for extracting information
NLP Cloud serves high performance pre-trained or custom models for NER
A fast, helpful, and open-source document parser
Parse text and tables from PDF files.
Document content and metadata extraction microservice
A cross-platform software for text translation and recognition
Open source NLP guide with models, methods, and real use cases
OCR software, free and offline
Browser utility for extracting and organizing webpage text quickly.
Python module for parsing semi-structured text into python tables
NLP Cloud serves high performance pre-trained or custom models for NER
A self-hostable bookmark-everything app
A Family of Open Sourced Music Foundation Models
Archive of leaked AI system prompts and internal instruction sets
Knowledge Graph Generation from Any Text