WebNews Crawler is a specific web crawler (spider, fetcher) designed to acquire and clean news articles from RSS and HTML pages. It can do a site specific extraction to extract the actual news content only, filtering out the advertising and other cruft.

Project Activity

See All Activity >

License

GNU General Public License version 2.0 (GPLv2)

Follow WebNews Crawler

WebNews Crawler Web Site

Other Useful Business Software
Outgrown Windows Task Scheduler? Icon
Outgrown Windows Task Scheduler?

Free diagnostic identifies where your workflow is breaking down—with instant analysis of your scheduling environment.

Windows Task Scheduler wasn't built for complex, cross-platform automation. Get a free diagnostic that shows exactly where things are failing and provides remediation recommendations. Interactive HTML report delivered in minutes.
Download Free Tool
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of WebNews Crawler!

Additional Project Details

Intended Audience

Advanced End Users

User Interface

Command-line

Programming Language

Java

Related Categories

Java Search Engines, Java Web Scrapers

Registered

2006-05-19