Using WebHarvy, you can easily scrape product details from Amazon Best Sellers list. Product data including name, price, ASIN, UPC, description, rating/reviews, specifications etc. can be extracted.
WebHarvy is a point-and-click web scraping software. It automatically identifies data patterns occurring in web pages. Product data can be extracted by following the product links from the Best Sellers list. Listings which span over multiple pages can be scraped by following the pagination links. WebHarvy also allows you to scrape images and download them to your local disk. The extracted data can be saved in various formats including CSV, Excel, XML, JSON etc.
Video Demonstration
Steps to be followed
Steps given below can be followed to scrape product details from Amazon Best Sellers list using WebHarvy:
- 1. Download and install WebHarvy in your computer
- 2. Open WebHarvy and load the Amazon Best Sellers list page
- 3. Start Configuration
- 4. Click and Select data from the listings page. Details like product name, thumbnail image URL, product page URL etc. can be selected from the listings page by clicking on them. During configuration, when you click over any data element on the page, a Capture window with various options will be displayed. Select the 'Capture Text' option to select the text of the clicked element for extraction.
- 5. Click on the next page link and select the Set as Next Page Link option from the Capture window.
- 6. To load the product details page, click on the first product link and select the Follow this link option.
- 7. Click and select the required product details from the product details page.
- 8. Stop Configuration. Save Configuration
- 9. Start Mine