Web data extraction (web data mining, web scraping) tool. It leverages well proved XML and text processing techologies in order to easely extract useful data from arbitrary web pages.

Project Activity

See All Activity >

License

BSD License, GNU General Public License version 2.0 (GPLv2)

Follow WebHarvest - web data extraction tool

WebHarvest - web data extraction tool Web Site

Other Useful Business Software
Outgrown Windows Task Scheduler? Icon
Outgrown Windows Task Scheduler?

Free diagnostic identifies where your workflow is breaking down—with instant analysis of your scheduling environment.

Windows Task Scheduler wasn't built for complex, cross-platform automation. Get a free diagnostic that shows exactly where things are failing and provides remediation recommendations. Interactive HTML report delivered in minutes.
Download Free Tool
Rate This Project
Login To Rate This Project

User Ratings

★★★★★
★★★★
★★★
★★
10
1
1
1
1
ease 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5
features 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5
design 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 3 / 5
support 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5

User Reviews

  • Yeah, it works well for web data extraction. But it is not enough powerful for cloud extraction. For this, I use another web scraping tool, octoparse.
  • I've used this tool several times on a dozen of so different sites with good results. The syntax can be challenging. Once you get used to it, it works quite well. Support was very good in the past. Very helpful. Sorry to see development has stopped by the looks of it.
  • Great use of XSLT and visual representation. Would be better to easier identify the results of the search. I prefer this htp://webminer.avantprime.com however for data extraction.
  • All other 18 reviews are FAKE and by the uploader.
  • dont find any donation button ...
    1 user found this review helpful.
Read more reviews >

Additional Project Details

Operating Systems

Linux

Intended Audience

Advanced End Users, Developers

User Interface

Java Swing

Programming Language

Java, XSL (XSLT/XPath/XSL-FO)

Database Environment

MySQL

Related Categories

XSL (XSLT/XPath/XSL-FO) XML Software, XSL (XSLT/XPath/XSL-FO) HTML XHTML, XSL (XSLT/XPath/XSL-FO) Search Engines, XSL (XSLT/XPath/XSL-FO) Frameworks, XSL (XSLT/XPath/XSL-FO) Web Scrapers, Java XML Software, Java HTML XHTML, Java Search Engines, Java Frameworks, Java Web Scrapers

Registered

2006-07-14