Web data extraction (web data mining, web scraping) tool. It leverages well proved XML and text processing techologies in order to easely extract useful data from arbitrary web pages.

Project Activity

See All Activity >

License

BSD License, GNU General Public License version 2.0 (GPLv2)

Follow WebHarvest - web data extraction tool

WebHarvest - web data extraction tool Web Site

Other Useful Business Software
Gain Advanced Threat Protection for Your AWS Workloads Icon
Gain Advanced Threat Protection for Your AWS Workloads

Running FortiGate NGFW on AWS Graviton2 Lets You Boost Scalability With Reduced Compute Costs

FortiGate-VM delivers comprehensive security and scalable VPN connectivity for your AWS workloads, while native AWS integrations unlock broad coverage for your environment. Now with support for AWS Graviton2 instances, FortiGate lets you optimize price performance and reduce your Amazon EC2 costs by up to 20 percent. Deploy today in AWS Marketplace.
Rate This Project
Login To Rate This Project

User Ratings

★★★★★
★★★★
★★★
★★
10
1
1
1
1
ease 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5
features 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5
design 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 3 / 5
support 1 of 5 2 of 5 3 of 5 4 of 5 5 of 5 2 / 5

User Reviews

  • Yeah, it works well for web data extraction. But it is not enough powerful for cloud extraction. For this, I use another web scraping tool, octoparse.
  • I've used this tool several times on a dozen of so different sites with good results. The syntax can be challenging. Once you get used to it, it works quite well. Support was very good in the past. Very helpful. Sorry to see development has stopped by the looks of it.
  • Great use of XSLT and visual representation. Would be better to easier identify the results of the search. I prefer this htp://webminer.avantprime.com however for data extraction.
  • All other 18 reviews are FAKE and by the uploader.
  • dont find any donation button ...
    1 user found this review helpful.
Read more reviews >

Additional Project Details

Operating Systems

Linux

Intended Audience

Advanced End Users, Developers

User Interface

Java Swing

Programming Language

XSL (XSLT/XPath/XSL-FO), Java

Database Environment

MySQL

Related Categories

XSL (XSLT/XPath/XSL-FO) XML Software, XSL (XSLT/XPath/XSL-FO) HTML XHTML, XSL (XSLT/XPath/XSL-FO) Search Engines, XSL (XSLT/XPath/XSL-FO) Frameworks, XSL (XSLT/XPath/XSL-FO) Web Scrapers, Java XML Software, Java HTML XHTML, Java Search Engines, Java Frameworks, Java Web Scrapers

Registered

2006-07-14