Ecommerce Product Page Scraper: Price, Stock, GTIN from URLs avatar

Ecommerce Product Page Scraper: Price, Stock, GTIN from URLs

Pricing

$1.50 / 1,000 product extracteds

Go to Apify Store
Ecommerce Product Page Scraper: Price, Stock, GTIN from URLs

Ecommerce Product Page Scraper: Price, Stock, GTIN from URLs

Paste product page URLs from almost any online store and get name, price, currency, sale price, stock status, brand, SKU, GTIN, rating, images and variants, read from the Schema.org and Open Graph data stores publish. Plain HTTP, no browser. Blocked pages are free.

Pricing

$1.50 / 1,000 product extracteds

Rating

0.0

(0)

Developer

Hay Equipos

Hay Equipos

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

2 days ago

Last modified

Share

Paste links to product pages from almost any online store and get one clean row per product: name, brand, price, currency, list price and sale flag, stock status, SKU, GTIN (EAN or UPC), MPN, rating, review count, images and every variant the page lists.

It works on Shopify, WooCommerce, BigCommerce, Magento, Wix, Squarespace, PrestaShop and custom stores, and on large retailers that let ordinary visitors in, because it reads the Schema.org product data (JSON-LD and microdata) and Open Graph product tags that stores publish for Google Shopping and social previews. No per site setup, no browser, no AI guessing.

What you can use it for

  • Price monitoring: schedule a list of competitor product pages daily and track price, sale price and stock.
  • Catalog enrichment: fill in GTINs, brands, images and descriptions for products you resell.
  • MAP and reseller checks: see what each reseller charges for the same GTIN.
  • Market research: compare prices, ratings and review counts across stores.
  • AI agents: one job, one input (a list of URLs), a predictable schema.

How it works

  1. Fetches each page with a plain HTTP request (one request at a time per site, with a pause between requests).
  2. Reads JSON-LD Product and ProductGroup blocks first (including @graph, AggregateOffer, list price specifications and hasVariant), then fills gaps from microdata, then from Open Graph product: tags.
  3. Pages that are blocked, missing or have no product data are written with a clear status and are not charged.

It respects robots.txt by default (including rules that name Apify), never logs in, never uses cookies from an account, never solves captchas and never uses proxies to get around blocks.

Input

FieldWhat it doesDefault
Product page URLsLinks to single product pages, one per linerequired
Include descriptionPlain text description, up to 5,000 characterson
Include variantsSize, color, SKU, GTIN, price and stock per varianton
Respect robots.txtSkip pages the site closes to automated toolson
Parallel requestsPages fetched at once across different sites5
Pause between requests to the same sitePoliteness delay in milliseconds1000
Maximum URLsStop after this many URLs1000

Example input:

{
"productUrls": [
"https://www.allbirds.com/products/mens-tree-runners",
"https://www.ikea.com/us/en/p/billy-bookcase-white-00263850/"
],
"includeDescription": true,
"includeVariants": true
}

Output

One row per URL. Example (trimmed):

{
"url": "https://www.allbirds.com/products/mens-tree-runners",
"status": "ok",
"dataSources": ["json-ld", "open-graph"],
"domain": "allbirds.com",
"name": "Men's Tree Runner",
"brand": "Allbirds",
"price": 100,
"currency": "USD",
"listPrice": null,
"onSale": null,
"lowPrice": 100,
"highPrice": 100,
"availability": "InStock",
"inStock": true,
"sku": "MENS_TREE_RUNNERS",
"gtin": null,
"rating": null,
"reviewCount": null,
"imageUrl": "https://cdn.shopify.com/s/files/1/1104/4168/files/...png",
"variantCount": 49,
"variants": [
{ "name": "Men's Tree Runner, Jet Black (White Sole), 8", "sku": "...", "gtin": "...", "size": "8", "price": 100, "currency": "USD", "availability": "InStock", "inStock": true }
],
"scrapedAt": "2026-09-27T12:00:00.000Z"
}

status is one of:

StatusMeaningCharged
okProduct data foundyes
no_product_dataPage loaded but has no usable product data (no price and no Schema.org product)no
blockedThe site answered with a bot check or access denied pageno
http_error404, 410, 500 and similarno
skipped_robots_txtrobots.txt closes the page to automated toolsno
errorTimeout or network failureno

A RUN_SUMMARY record in the run's key value store counts each status.

Pricing

Pay per event. No subscription and no platform usage charge on top.

EventPrice
Product extracted$0.0015 (1.50 dollars per 1,000 products)

Blocked pages, pages without product data, errors and robots.txt skips are free. Example: tracking 500 competitor products every day costs about $0.75 a day. Set a maximum charge per run in Apify and the actor stops cleanly when it is reached.

Limits

  • Sites with strong bot protection (for example Amazon, Walmart, Best Buy, Target, Home Depot and others using Akamai, PerimeterX, DataDome or Cloudflare challenges) often answer datacenter traffic with a block page. Those rows come back as blocked and cost nothing. This actor does not try to get around blocks.
  • Only what the page publishes as structured data is returned. If a store leaves out GTIN or stock in its Schema.org data, those fields are empty.
  • Prices are the ones shown to a visitor from a US datacenter with no cookies. Stores that change price or currency by country may show different values to you.
  • Prices rendered only by JavaScript after the page loads, with no structured data, are not seen (no browser is used).
  • One product per URL. Category and search pages are not crawled; use a sitemap or catalog tool to collect product URLs first.

FAQ

Which stores work best? Shopify, WooCommerce, BigCommerce, Wix, Squarespace and most modern stores publish complete Schema.org product data because Google Shopping needs it.

Do I pay for pages that fail? No. Only rows with status ok are charged.

Can I use it for price alerts? Yes. Schedule it, then compare price, listPrice and inStock between runs in your sheet, database or automation tool.

Does it use AI? No. It reads the data the store itself publishes, so results are exact and repeatable.

Is it affiliated with any store? No. It is an independent tool that reads public product pages.