Web data extraction (web data mining, web scraping) tool. It leverages well proved XML and text processing techologies in order to easely extract useful data from arbitrary web pages.

Project Activity

See All Activity >

License

BSD License, GNU General Public License version 2.0 (GPLv2)

Follow WebHarvest - web data extraction tool

WebHarvest - web data extraction tool Web Site

Other Useful Business Software
MongoDB Atlas runs apps anywhere Icon
MongoDB Atlas runs apps anywhere

Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Rate This Project
Login To Rate This Project

User Ratings

★★★★★
★★★★
★★★
★★
10
1
1
1
1
ease 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5
features 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5
design 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 3 / 5
support 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5

User Reviews

  • Yeah, it works well for web data extraction. But it is not enough powerful for cloud extraction. For this, I use another web scraping tool, octoparse.
  • I've used this tool several times on a dozen of so different sites with good results. The syntax can be challenging. Once you get used to it, it works quite well. Support was very good in the past. Very helpful. Sorry to see development has stopped by the looks of it.
  • Great use of XSLT and visual representation. Would be better to easier identify the results of the search. I prefer this htp://webminer.avantprime.com however for data extraction.
  • All other 18 reviews are FAKE and by the uploader.
  • dont find any donation button ...
    1 user found this review helpful.
Read more reviews >

Additional Project Details

Operating Systems

Linux

Intended Audience

Advanced End Users, Developers

User Interface

Java Swing

Programming Language

Java, XSL (XSLT/XPath/XSL-FO)

Database Environment

MySQL

Related Categories

XSL (XSLT/XPath/XSL-FO) XML Software, XSL (XSLT/XPath/XSL-FO) HTML XHTML, XSL (XSLT/XPath/XSL-FO) Search Engines, XSL (XSLT/XPath/XSL-FO) Frameworks, XSL (XSLT/XPath/XSL-FO) Web Scrapers, Java XML Software, Java HTML XHTML, Java Search Engines, Java Frameworks, Java Web Scrapers

Registered

2006-07-14