Web Product Scraper

Responsible web scraping:
permissions, pauses and privacy.

Collecting product data from public shop pages is everyday work for resellers and buyers, and it can be done in a way that is fair to the shop and safe for you. This page explains what Web Product Scraper does by default to stay polite, what it does not do for you, and the questions to ask before you reuse what you collect.

Add to Chrome, it's free No account · no server · nothing leaves your browser
Follow Item Links settings in Web Product Scraper: product pages opened a few at a time with a pause, pictures blocked in those tabs

What is responsible web scraping?

Taking only what you need from pages that are public, at a pace the site barely notices, stopping when it says no, and respecting what its terms and copyright allow you to do with the data. A tool can handle the pace; the decision about use stays with you.

What the extension does to stay polite

A pause between pages

3 seconds by default, adjustable from 1 to 60, plus a random extra of up to half that, so the visits do not arrive in a fixed rhythm.

A few tabs at a time

Product pages open 3 at a time by default, never more than 10, in one collapsed tab group that closes when the run ends.

Lighter visits

Pictures are blocked in the tabs it opens to read product pages, so the shop serves less; the full-size photos are downloaded when you ask for the ZIP.

Stops when refused

After three failed pages in a row the run stops by itself, instead of retrying against a shop that is turning it away.

Fewer requests on Shopify

Shopify collections come from the store's public product feed, up to 250 products per request, instead of one visit per product.

Only pages you open

A run starts from the page in your tab and the pages it leads to. There is no background crawling and no list of sites.

Permissions and your data

Asked per site

Chrome's permission is requested for one website, when you start working there, never for every site at once.

No server

The extension talks to no server of its own; what you collect stays in your browser until you export or clear it.

Settings follow your Chrome

Settings and taught rules are saved with Chrome's sync storage, so they appear on your other computers signed in to Chrome.

What it does not do for you

Web Product Scraper does not read a site's robots.txt file or its terms of use. It reads only the pages you open yourself and the pages a run follows from them, so the choice of site is yours, and so is checking whether that site allows automated collection.

If a shop's terms forbid it, or the shop asks you to stop, stop. Lowering the pace does not change what the terms say.

Using what you collect

Terms of use

Each site sets its own rules for automated access and reuse. Read them before a large run.

Photos and descriptions

They are protected by copyright. Supplier material is often meant for resellers; ask before you publish it anywhere else.

Personal data

Collect products, not people. Names, emails or reviews of individuals need a lawful reason to keep, and this tool is not built for them.

Prices

Comparing public prices is common practice; republishing a competitor's whole catalogue is not.

A short checklist before a run

Read the terms

Look for rules on automated access and on reusing content.

Take only what you need

Set the page range and Max products to the job, not to the whole shop.

Keep the default pace

Three pages at a time with a few seconds' pause is gentle on most shops.

Ask about photos

Before listing a supplier's products, check that you may use their pictures.

Frequently asked questions

Is web scraping legal?

It depends on the site's terms, the data and where you are; this page is not legal advice. Reading public product pages at a gentle pace is common, but you are responsible for how you use what you collect.

Does Web Product Scraper respect robots.txt?

It does not read robots.txt. It only visits pages you open and the pages a run follows from them; check a site's rules yourself before a large run.

Can a shop tell that I am scraping?

The visits come from your own browser, but a shop can notice many pages opened quickly from one address and block it, which is why the pause and the stop on refusal exist.

Does the extension send my data anywhere?

No. It has no server of its own; your list stays in your browser until you export or clear it.

Add to Chrome, it's free