A fast, high-level web crawling and web scraping framework
Python crawler for collecting and downloading Sina Weibo user data
Open-Source RPA Software (formerly Kantu)
Cross platform GUI tool for downloading videos from Bilibili sites
Python tool for crawling and extracting structured data from news site
Python crawler to download photos and videos from Tumblr blogs
Realtime crawler for COVID-19 outbreak statistics from DXY data
Lightweight Python tool for downloading videos from many platforms
A scalable web crawler framework for Java
CLI tool to save complete web pages as single self-contained HTML file
Scrape job websites into a single spreadsheet with no duplicates.
The unix-way web crawler
High-performance Rust web crawler and scraper for large-scale data
Asyncio-based Python framework for building fast web crawling spiders
Open source web scraping system for automated data collection tasks
Lighter, faster browser kernel of blink to integrate HTML UI in apps
Fast CLI web crawler for discovering endpoints in modern web apps
Remote isolated browser API for security
Fast CLI tool for cloning entire websites for local browsing offline
Collection of JS reverse engineering examples for web scraping study
Easily turn large sets of image urls to an image dataset
Web Scraper in Go, similar to BeautifulSoup
Python & command-line tool to gather text on the Web
Open source Douyin crawler for collecting and downloading public data
Python crawler and API for downloading JMComic albums and images