Yozh Scraper
Yozh Scraper is a powerful open-source web scraping and crawling toolkit built for high-scale data extraction. Powered by Playwright, Python, and Redis, it effortlessly handles complex JS-rendered sites while bypassing modern anti-bot protections.
Key Capabilities:
• Anti-Detect Scraping: Leverages Camoufox and real Chrome instances to spoof browser fingerprints and overcome strict anti-scraping systems.
• Dual Microservices: Includes an async Scraper API for page rendering and an Open Crawler with SSE streaming, site-mapping, and deduplication.
• Native MCP Support: Directly integrates with AI agents (Claude Code/Desktop, LangChain, n8n) via built-in Model Context Protocol (/mcp) endpoints.
• Smart Parsing & Presets: Pre-configured for Amazon, Google, LinkedIn, and more, featuring optional LLM self-healing parsing.
• Enterprise Scaling: Horizontal worker scaling via Docker Compose, proxy support (Residential/Mobile/DC), and a web UI for testing.
Learn more
Decodo
Decodo (formerly Smartproxy) offers advanced proxy infrastructure and web scraping solutions to streamline web data collection for businesses and developers. With over 125 million ethically sourced IP addresses (residential, mobile, datacenter, and static residential proxies), Decodo helps users efficiently bypass geo-restrictions, CAPTCHAs, and other web access barriers. Decodo's intuitive APIs enable effortless, structured data scraping from websites, eCommerce platforms, search engines, and social media, supporting outputs in HTML, JSON, and CSV formats. The platform includes the Universal Scraper for easy real-time data extraction and an upcoming AI-powered Parser to minimize tedious manual data processing. Ideal for price aggregation, SEO monitoring, ad verification, multi-account management, AI training, and private browsing. Decodo also offers comprehensive documentation, responsive support, and transparent policies, including a 3-day trial and clear refund guidelines.
Learn more
ScrapeBadger
ScrapeBadger is a web scraping API platform specialising in Twitter/X, Reddit and Google data, with dedicated scrapers also covering TikTok, YouTube, LinkedIn, Amazon, eBay, Zillow and 40+ more: with built-in anti-bot bypass and an MCP server for AI agents.
Handles Cloudflare, DataDome, Akamai, Imperva, PerimeterX, and Kasada automatically. No proxy management, no CAPTCHA solving, no broken scripts. Every scraper returns clean structured JSON. Failed requests are never charged.
Covers social media, Google (18 products), e-commerce (Amazon, eBay, Vinted, Leboncoin, Depop), and real estate (Zillow, Redfin, Realtor, Idealista, Immobiliare, LoopNet). MCP server connects all scrapers to AI agents including Claude, ChatGPT, and Cursor. Official Node.js and Python SDKs.
Learn more
WebscrapeAi
WebscrapeAi is the perfect tool for collecting data from the web without the hassle of manual scraping. No coding skills are required. Simply enter the URL and the items you want to scrape, and our AI scraper will do the rest. Our AI scraper uses advanced algorithms to collect data accurately, so you can be confident in the results. With our AI scraper, you can automate your data collection process and free up your time to focus on other tasks. Our AI scraper allows you to easily customize your data collection preferences to suit your needs. Our AI scraper is an affordable solution for businesses of all sizes that want to collect data without breaking the bank. Our AI scraper uses state-of-the-art methods for data collection to ensure speedy collection of data. An AI scraper is a tool that uses artificial intelligence algorithms to automatically collect data from websites. It's legal to use an AI scraper to collect publicly available data.
Learn more