The easiest way to parse and modify URLs in Python
Web Scraping Framework
Open source file indexing & storage analytics powered by Elasticsearch
Redis-based components for Scrapy
Multiprocess Selenium crawler for downloading images by keywords
dude uncomplicated data extraction: A simple framework
Command-line Bilibili video and danmaku downloader with batch support
Web crawler that finds hidden web directories without brute force
Distributed Crawler Management Framework Based on Scrapy
Distributed web crawler admin platform for spiders management
A service daemon to run Scrapy spiders
Python library providing APIs for automated website login workflows
Web crawler for archiving and backing up sites into WARC archives
A Smart, Automatic, Fast and Lightweight Web Scraper for Python
ML-based HTML scraper that learns extraction rules from examples
Intelligent proxy pool for collecting and managing public proxies
Instagram profile crawler that extracts posts, tags, and stats
Async Python framework for fast and flexible web scraping spiders
Python tool for scraping search engine results from many providers
Collection of Python ecommerce and website crawler examples projects
Asynchronous tool for finding and checking public proxy servers
Pythonic HTML Parsing for Humans
Creating Scrapy scrapers via the Django admin interface
Python crawler that downloads image galleries and analyzes titles