Document (PDF, Word, PPTX ...) extraction and parse API
Recognition and resolution of numbers, units, date/time, etc.
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
JavaScript OCR and text extraction for images and PDFs
Python & command-line tool to gather text on the Web
Zero-copy PDF text extraction library written in Zig
CLI tool to extract (meta)data from PDF and manipulate PDF files
MD/.JSON Document OCR and structured data extraction API
Official Vectorize MCP Server
Structured data extraction and instruction calling with ML, LLM
Fast and efficient unstructured data extraction
NLP Cloud serves high performance pre-trained or custom models for NER
A machine learning software for extracting information
A fast, helpful, and open-source document parser
Parse text and tables from PDF files.
A cross-platform software for text translation and recognition
Document content and metadata extraction microservice
Open source NLP guide with models, methods, and real use cases
OCR software, free and offline
Python module for parsing semi-structured text into python tables
Browser utility for extracting and organizing webpage text quickly.
NLP Cloud serves high performance pre-trained or custom models for NER
A Family of Open Sourced Music Foundation Models
Archive of leaked AI system prompts and internal instruction sets
A self-hostable bookmark-everything app