A fast, high-level web crawling and web scraping framework
Python crawler for collecting and downloading Sina Weibo user data
Open-Source RPA Software (formerly Kantu)
Cross platform GUI tool for downloading videos from Bilibili sites
Python tool for crawling and extracting structured data from news site
Python crawler to download photos and videos from Tumblr blogs
Realtime crawler for COVID-19 outbreak statistics from DXY data
Lightweight Python tool for downloading videos from many platforms
CLI tool to save complete web pages as single self-contained HTML file
A scalable web crawler framework for Java
Scrape job websites into a single spreadsheet with no duplicates.
Asyncio-based Python framework for building fast web crawling spiders
High-performance Rust web crawler and scraper for large-scale data
Fast CLI web crawler for discovering endpoints in modern web apps
Open source web scraping system for automated data collection tasks
Lighter, faster browser kernel of blink to integrate HTML UI in apps
Remote isolated browser API for security
A Node.js scraper for humans
Fast CLI tool for cloning entire websites for local browsing offline
Collection of JS reverse engineering examples for web scraping study
Web Scraper in Go, similar to BeautifulSoup
Easily turn large sets of image urls to an image dataset
Open source Douyin crawler for collecting and downloading public data
Python & command-line tool to gather text on the Web
Python crawler and API for downloading JMComic albums and images