Turn entire websites into LLM-ready markdown or structured data
High-performance Rust web crawler and scraper for large-scale data
AI-ready web crawler that extracts and structures website content
A fast, high-level web crawling and web scraping framework
Fast CLI web crawler for discovering endpoints in modern web apps
Python crawler to download photos and videos from Tumblr blogs
Blazing fast Go framework for web crawling and data scraping tasks
The unix-way web crawler
Python tool for crawling and extracting structured data from news site
Library for Rapid (Web) Crawler and Scraper Development
The complete web scraping toolkit for PHP
Free batch downloader for image, wallpaper, video, audio, document,
Distributed Crawler Management Framework Based on Scrapy
Goutte, a simple PHP Web Scraper
Web crawler for archiving and backing up sites into WARC archives
Instagram profile crawler that extracts posts, tags, and stats
Fast and flexible C# framework for building customizable web crawlers
Gospider - Fast web spider written in Go
Polite concurrent web crawler library for Go with flexible hooks
The next web scraper, see through the <html> noise
Pixiv crawler userscript for downloading artwork and galleries easily
Open source web crawler for Java
A powerful Spider(Web Crawler) system in Python
Python library to crawl and retrieve data from WeChat accounts
Website crawler that audits site pages automatically with Lighthouse