WebNews Crawler is a specific web crawler (spider, fetcher) designed to acquire and clean news articles from RSS and HTML pages. It can do a site specific extraction to extract the actual news content only, filtering out the advertising and other cruft.

Project Activity

See All Activity >

License

GNU General Public License version 2.0 (GPLv2)

Follow WebNews Crawler

WebNews Crawler Web Site

Other Useful Business Software
Go from Code to Production URL in Seconds Icon
Go from Code to Production URL in Seconds

Cloud Run deploys apps in any language instantly. Scales to zero. Pay only when code runs.

Skip the Kubernetes configs. Cloud Run handles HTTPS, scaling, and infrastructure automatically. Two million requests free per month.
Try it free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of WebNews Crawler!

Additional Project Details

Intended Audience

Advanced End Users

User Interface

Command-line

Programming Language

Java

Related Categories

Java Search Engines, Java Web Scrapers

Registered

2006-05-19