Ecommerce Product Page Scraper: Price, Stock, GTIN from URLs
Pricing
$1.50 / 1,000 product extracteds
Ecommerce Product Page Scraper: Price, Stock, GTIN from URLs
Paste product page URLs from almost any online store and get name, price, currency, sale price, stock status, brand, SKU, GTIN, rating, images and variants, read from the Schema.org and Open Graph data stores publish. Plain HTTP, no browser. Blocked pages are free.
Pricing
$1.50 / 1,000 product extracteds
Rating
0.0
(0)
Developer
Hay Equipos
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
2 days ago
Last modified
Categories
Share
Paste links to product pages from almost any online store and get one clean row per product: name, brand, price, currency, list price and sale flag, stock status, SKU, GTIN (EAN or UPC), MPN, rating, review count, images and every variant the page lists.
It works on Shopify, WooCommerce, BigCommerce, Magento, Wix, Squarespace, PrestaShop and custom stores, and on large retailers that let ordinary visitors in, because it reads the Schema.org product data (JSON-LD and microdata) and Open Graph product tags that stores publish for Google Shopping and social previews. No per site setup, no browser, no AI guessing.
What you can use it for
- Price monitoring: schedule a list of competitor product pages daily and track price, sale price and stock.
- Catalog enrichment: fill in GTINs, brands, images and descriptions for products you resell.
- MAP and reseller checks: see what each reseller charges for the same GTIN.
- Market research: compare prices, ratings and review counts across stores.
- AI agents: one job, one input (a list of URLs), a predictable schema.
How it works
- Fetches each page with a plain HTTP request (one request at a time per site, with a pause between requests).
- Reads
JSON-LDProductandProductGroupblocks first (including@graph,AggregateOffer, list price specifications andhasVariant), then fills gaps from microdata, then from Open Graphproduct:tags. - Pages that are blocked, missing or have no product data are written with a clear
statusand are not charged.
It respects robots.txt by default (including rules that name Apify), never logs in, never uses cookies from an account, never solves captchas and never uses proxies to get around blocks.
Input
| Field | What it does | Default |
|---|---|---|
| Product page URLs | Links to single product pages, one per line | required |
| Include description | Plain text description, up to 5,000 characters | on |
| Include variants | Size, color, SKU, GTIN, price and stock per variant | on |
| Respect robots.txt | Skip pages the site closes to automated tools | on |
| Parallel requests | Pages fetched at once across different sites | 5 |
| Pause between requests to the same site | Politeness delay in milliseconds | 1000 |
| Maximum URLs | Stop after this many URLs | 1000 |
Example input:
{"productUrls": ["https://www.allbirds.com/products/mens-tree-runners","https://www.ikea.com/us/en/p/billy-bookcase-white-00263850/"],"includeDescription": true,"includeVariants": true}
Output
One row per URL. Example (trimmed):
{"url": "https://www.allbirds.com/products/mens-tree-runners","status": "ok","dataSources": ["json-ld", "open-graph"],"domain": "allbirds.com","name": "Men's Tree Runner","brand": "Allbirds","price": 100,"currency": "USD","listPrice": null,"onSale": null,"lowPrice": 100,"highPrice": 100,"availability": "InStock","inStock": true,"sku": "MENS_TREE_RUNNERS","gtin": null,"rating": null,"reviewCount": null,"imageUrl": "https://cdn.shopify.com/s/files/1/1104/4168/files/...png","variantCount": 49,"variants": [{ "name": "Men's Tree Runner, Jet Black (White Sole), 8", "sku": "...", "gtin": "...", "size": "8", "price": 100, "currency": "USD", "availability": "InStock", "inStock": true }],"scrapedAt": "2026-09-27T12:00:00.000Z"}
status is one of:
| Status | Meaning | Charged |
|---|---|---|
ok | Product data found | yes |
no_product_data | Page loaded but has no usable product data (no price and no Schema.org product) | no |
blocked | The site answered with a bot check or access denied page | no |
http_error | 404, 410, 500 and similar | no |
skipped_robots_txt | robots.txt closes the page to automated tools | no |
error | Timeout or network failure | no |
A RUN_SUMMARY record in the run's key value store counts each status.
Pricing
Pay per event. No subscription and no platform usage charge on top.
| Event | Price |
|---|---|
| Product extracted | $0.0015 (1.50 dollars per 1,000 products) |
Blocked pages, pages without product data, errors and robots.txt skips are free. Example: tracking 500 competitor products every day costs about $0.75 a day. Set a maximum charge per run in Apify and the actor stops cleanly when it is reached.
Limits
- Sites with strong bot protection (for example Amazon, Walmart, Best Buy, Target, Home Depot and others using Akamai, PerimeterX, DataDome or Cloudflare challenges) often answer datacenter traffic with a block page. Those rows come back as
blockedand cost nothing. This actor does not try to get around blocks. - Only what the page publishes as structured data is returned. If a store leaves out GTIN or stock in its Schema.org data, those fields are empty.
- Prices are the ones shown to a visitor from a US datacenter with no cookies. Stores that change price or currency by country may show different values to you.
- Prices rendered only by JavaScript after the page loads, with no structured data, are not seen (no browser is used).
- One product per URL. Category and search pages are not crawled; use a sitemap or catalog tool to collect product URLs first.
FAQ
Which stores work best? Shopify, WooCommerce, BigCommerce, Wix, Squarespace and most modern stores publish complete Schema.org product data because Google Shopping needs it.
Do I pay for pages that fail? No. Only rows with status ok are charged.
Can I use it for price alerts? Yes. Schedule it, then compare price, listPrice and inStock between runs in your sheet, database or automation tool.
Does it use AI? No. It reads the data the store itself publishes, so results are exact and repeatable.
Is it affiliated with any store? No. It is an independent tool that reads public product pages.