Jaunt
Jaunt is a Java library designed for web scraping, web automation, and JSON querying. It provides a fast, ultra-light headless browser that enables Java programs to perform tasks such as web scraping, form handling, and interfacing with REST APIs. Jaunt supports parsing of HTML, XHTML, XML, and JSON, and offers features like HTTP header and cookie manipulation, proxy support, and customizable caching. The library does not support JavaScript execution; however, for automating JavaScript-enabled browsers, Jauntium is recommended. Jaunt is available under the Apache License, with a monthly edition that expires periodically, requiring users to download the latest version upon expiration. The library is suitable for tasks such as parsing and extracting data from web pages, filling out and submitting forms, and handling HTTP requests and responses. Comprehensive tutorials and documentation are available to assist users in getting started with Jaunt.
Learn more
Crawlora
Crawlora is a structured web-data platform. Instead of building and maintaining scrapers, you call documented REST endpoints — or 319 hosted MCP tools — and get normalized JSON instead of HTML to parse. It spans 393 endpoints across search/SERP (Google, Bing, Brave), maps, e-commerce (Amazon, eBay, Shopify), app stores, social (TikTok, YouTube, Instagram, Reddit), reviews, and finance. Crawlora handles proxy rotation, headless-browser rendering, and retries behind the API, so your team ships data features instead of operating scraping infrastructure. The same endpoints are exposed as a Model Context Protocol (MCP) server, so AI agents in Claude, Cursor, Cline, or n8n pull live web data with one header. Pricing is pay-on-success — billed only on a successful (2xx) response — with a free tier of 2,000 credits/month (no card) and a public Playground to run any endpoint and see the JSON before writing code.
Learn more
Social Fetch
Social Fetch is a REST API that provides synchronous access to structured JSON for public posts, profiles, comments, engagement metrics, transcripts, and search results across Instagram, TikTok, X/Twitter, YouTube, Facebook, LinkedIn, Reddit, Threads, Telegram, GitHub, Spotify, Rumble, and general web pages. It abstracts all scraping infrastructure, proxies, headless browsers, DOM changes, and rate limits, returning a unified cross-platform schema so developers can add or switch platforms without rewriting client logic.
Learn more
PYPROXY
Market-leading proxy solution provides tens of millions of IP resources. Commercial residential and ISP proxy network includes 90M+ IPs around the world. Exclusive high-performance server requests access to real residential addresses. Abundant bandwidth support business demands. Real-time speed can reach 1M-5M/s.99% success rate guarantee data collection activities. There is no limit to the number of uses or invocation frequencies of the proxies. You can generate huge amounts of proxies at one time. Provide various API parameter configurations. Generate proxies by the method of username & password authentication, convenient and fast. Get highly anonymous real residential IPs and your privacy and safety are completely protected. Your real network environment won't be acquired at any time. Exclusive high-performance server requests access in real residential addresses, maintaining the normal connection of the proxies. Unlimited concurrency reduces business costs.
Learn more