WebNews Crawler is a specific web crawler (spider, fetcher) designed to acquire and clean news articles from RSS and HTML pages. It can do a site specific extraction to extract the actual news content only, filtering out the advertising and other cruft.

Project Activity

See All Activity >

License

GNU General Public License version 2.0 (GPLv2)

Follow WebNews Crawler

WebNews Crawler Web Site

You Might Also Like
Omnichannel contact center platform for enterprises. Icon
Omnichannel contact center platform for enterprises.

For Call centers or BPOs with a very high volume of calls

Deliver a personalized customer experience with every interaction, across every channel, with uContact, net2phone’s cloud contact center solution.
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of WebNews Crawler!

Additional Project Details

Intended Audience

Advanced End Users

User Interface

Command-line

Programming Language

Java

Related Categories

Java Search Engines, Java Web Scrapers

Registered

2006-05-19