Pandas on AWS, easy integration with Athena, Glue, Redshift, etc.
Ansible role which installs and configures Graylog
Powerful Python crawler framework for scalable web scraping tasks
Changelog CI is a GitHub Action that enables a project
Python & command-line tool to gather text on the Web
Dataproc templates and pipelines for solving simple in-cloud data task
Consolidate and extend hosts files from several well-curated sources
Asyncio-based Python framework for building fast web crawling spiders
Prevent cloud misconfigurations during build-time for Terraform
NBA Stats API via Basketball Reference
Fast and modern gateway for VLESS tunneling on WebSocket and XHTTP
FastAPI server-side rendering with built-in HTMX support.
Utilize all available CPU cores for accepting new client connections
Twitter for Python
Anti-Detect Browser that passes every bot detection test
Collection of JS reverse engineering examples for web scraping study
A library that scrapes Linkedin for user data
Ajenti Core and stock plugins
Movie metadata scraper and organizer for media libraries and NFO
Easily turn large sets of image urls to an image dataset
Factorio headless server in a Docker container
Easy-to-use and developer-friendly enterprise CMS powered by Django
The best free open source website change detection and restock service
Web Scraping Framework
NeoDB is a self-hosted server tracking what you read/watch/listen/play