ExtractThinker is a Document Intelligence library for LLMs
Structured data extraction and instruction calling with ML, LLM
Fast, local-first web content extraction for LLMs
No-code LLM Platform to launch APIs and ETL Pipelines
Crawl a website starting from a URL, find relevant pages
Model Context Protocol server that integrates AgentQL's data
Fast and efficient unstructured data extraction
ContextGem: Effortless LLM extraction from documents
Extract and convert data from any document, images, pdfs, word doc
A high-quality tool for convert PDF to Markdown and JSON
Document content and metadata extraction microservice
Document (PDF, Word, PPTX ...) extraction and parse API
A fast, helpful, and open-source document parser
RAGFlow is an open-source RAG (Retrieval-Augmented Generation) engine
Synthetic data curation for post-training and data extraction
Official Vectorize MCP Server
A powerful Model Context Protocol (MCP) server
Skill for installing full networking capabilities for Claude Code
Enhance any agent's browser use skill
Vision AI browser agent for automation, testing, and extraction
Claude Code skill for generating production-quality SVG+PNG technical
AI Browser Agent is an advanced Browser AI tool
End-to-end pipeline converting generative videos
Superlinked is a Python framework for AI Engineers
An on-premises, OCR-free unstructured data extraction