# Shopify Store Scraper (`brightpath-data/shopify-store-scraper`) Actor

Products, variants, prices, images and collections for any list of Shopify store domains, no contact data

- **URL**: https://apify.com/brightpath-data/shopify-store-scraper.md
- **Developed by:** [Nick Randall](https://apify.com/brightpath-data) (community)
- **Categories:** E-commerce, Other
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.00 / 1,000 product records

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Shopify Store Scraper

Get every product, variant, price and image from any Shopify store's public catalog, no login, no API key, no per-store setup fee that paid Shopify scraper Actors on the Store charge $3 to $10 per 1,000 results for.

Get clean, structured product data from any Shopify store's storefront as JSON, CSV or Excel, or call it as a tool from Claude, Cursor, ChatGPT or any MCP client. Pay only for the results you receive.

### What you get

For every domain you list, every product currently in the store's catalog: title, handle, vendor, product type, tags, timestamps, a canonical product URL, every variant (price, compare-at price, SKU, availability, weight, inventory policy) and every image (source URL, alt text, dimensions). Pulled straight from Shopify's own public `products.json` storefront endpoint, the same data your browser gets when you load a collection page.

Every result is a flat record with stable field names, so it drops straight into a spreadsheet, a database or an AI agent's context.

### Why use this instead of manual browsing or a browser-based scraper

- No login, no API key, no app install on the target store: this reads the same public JSON Shopify already serves to any visitor's browser
- No headless browser needed, so runs are fast and cheap compared to scrapers that render each product page
- Full variant and pricing detail in one pass, not just what is visible on a collection grid
- Results in JSON, CSV, Excel or via API, or piped into Zapier, Make, n8n and Google Sheets
- Works as an MCP tool, so AI agents can pull competitor pricing or catalog data on demand

### Input

| Field | Type | Default | Meaning |
|-------|------|---------|---------|
| `domains` | array | | Shopify store domains to scrape, e.g. `allbirds.com` or `https://allbirds.com`. Either form works; both are normalized to the same host. |
| `maxResults` | integer | 100 | Cap on results saved. You are charged per result, so this caps your cost. |

Example input:

```json
{
  "domains": ["allbirds.com", "gymshark.com"],
  "maxResults": 100
}
```

### Output

Example result:

```json
{
  "domain": "allbirds.com",
  "id": 7218356060240,
  "title": "Men's Strider Explore - Natural Black (Dark Grey Sole)",
  "handle": "mens-strider-explore",
  "url": "https://allbirds.com/products/mens-strider-explore",
  "vendor": "Allbirds",
  "productType": "Shoes",
  "tags": ["allbirds::edition => classic", "allbirds::gender => mens", "allbirds::silhouette => runner", "loop::returnable => true", "shoprunner"],
  "createdAt": "2025-09-10T12:47:47-07:00",
  "updatedAt": "2026-09-24T15:50:10-07:00",
  "publishedAt": "2026-08-26T10:22:14-07:00",
  "variants": [
    { "id": 41334293889104, "title": "8", "price": "130.00", "compareAtPrice": null, "sku": "A11768M080", "available": false, "weight": null, "inventoryPolicy": null },
    { "id": 41334293921872, "title": "8.5", "price": "130.00", "compareAtPrice": null, "sku": "A11768M085", "available": false, "weight": null, "inventoryPolicy": null }
  ],
  "images": [
    { "src": "https://cdn.shopify.com/s/files/1/1104/4168/files/A11768_25Q4_Strider-Explore-Natural-Black-Dark-Grey-Sole_PDP_LEFT.png?v=1759336475", "alt": null, "width": 4000, "height": 4000 }
  ],
  "optionsSummary": "Size"
}
```

Sample rows from the same live run (25 products from allbirds.com at `maxResults: 25`):

| domain | title | vendor | variants |
|---|---|---|---|
| allbirds.com | Free Returns Coverage | re:do | 2 |
| allbirds.com | Men's Strider Explore - Natural Black (Dark Grey Sole) | Allbirds | 13 |
| allbirds.com | Women's Dasher NZ - Blizzard/Deep Navy (Blizzard Sole) | Allbirds | 13 |
| allbirds.com | Men's Merino Blend Sweatpant - True Black | Allbirds | 5 |

Field reference:

- `domain` - the store's host, e.g. `allbirds.com`
- `id` - Shopify's numeric product ID
- `title`, `handle` - product title and URL slug
- `url` - canonical product page URL, built from `domain` and `handle`
- `vendor`, `productType` - as set by the store
- `tags` - array of tags (split from Shopify's comma-separated tag string)
- `createdAt`, `updatedAt`, `publishedAt` - ISO timestamps from the store
- `variants` - array of `{ id, title, price, compareAtPrice, sku, available, weight, inventoryPolicy }`
- `images` - array of `{ src, alt, width, height }`
- `optionsSummary` - the product's option names joined together, e.g. `"Size, Color"`

### Pricing

Pay per event. You are charged **$2.00 per 1,000 results** saved to the dataset, plus a fraction of a cent per run start. Nothing is charged for results you do not receive. Set "Max total charge per run" in the run options to cap spending on any run. When a run reaches your cap it stops cleanly and keeps everything it already saved.

Rough guide: 1,000 results cost $2.00 and take under two minutes.

### Use it from an AI agent (MCP)

This Actor is available as an MCP tool through the Apify MCP server. Add it to your client, then ask the agent for the data in plain language.

Claude Desktop, Claude Code or Cursor (`mcp.json` / `claude_desktop_config.json`):

```json
{
  "mcpServers": {
    "apify": {
      "url": "https://mcp.apify.com/?actors=brightpath-data/shopify-store-scraper",
      "headers": { "Authorization": "Bearer YOUR_APIFY_TOKEN" }
    }
  }
}
```

ChatGPT and other clients that support remote MCP servers: add `https://mcp.apify.com/?actors=brightpath-data/shopify-store-scraper` as a connector with your Apify token.

Example prompt once connected: "Pull every product and price from allbirds.com and gymshark.com so I can compare their shoe lineups."

### Use it from code

```bash
curl -X POST "https://api.apify.com/v2/acts/brightpath-data~shopify-store-scraper/run-sync-get-dataset-items?token=YOUR_APIFY_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"domains":["allbirds.com","gymshark.com"],"maxResults":100}'
```

Python:

```python
from apify_client import ApifyClient
client = ApifyClient("YOUR_APIFY_TOKEN")
run = client.actor("brightpath-data/shopify-store-scraper").call(run_input={"domains": ["allbirds.com", "gymshark.com"], "maxResults": 100})
items = client.dataset(run["defaultDatasetId"]).list_items().items
```

### Limits and fair use

- Up to 10,000 results per run, across as many domains as you list.
- Works on essentially any live store built on Shopify that has not specifically blocked its public `products.json` endpoint. A small number of stores rate-limit or block this endpoint; if a domain's first page returns a 404 or 403, that domain is skipped (logged as a warning) and the run continues with the rest.
- Requests are paced at least 500ms apart per the run, out of respect for stores that rate-limit bots.
- Only products currently published to the storefront are returned; draft or unpublished products are not visible through this endpoint and are not included.

### Data source and legal

This Actor reads only Shopify's public, unauthenticated storefront `products.json` endpoint, the same catalog JSON a store already serves to any visitor's browser when its collection pages load. It collects public product catalog data only: no customer data, no contact information, no order or checkout data is ever touched or collected. No login, account, app installation or access control is bypassed. This Actor collects public, non-personal data only and does not bypass logins, paywalls or access controls. You are responsible for how you use the data.

### Support

Found a problem or need a field added? Open an issue on the Actor's Issues tab. Fixes for broken runs are prioritized.

# Actor input Schema

## `domains` (type: `array`):

Shopify store domains to scrape, e.g. allbirds.com or https://allbirds.com

## `maxResults` (type: `integer`):

Maximum number of products to save. You are charged per result saved, so this also caps the cost of a run.

## Actor input object example

```json
{
  "domains": [
    "allbirds.com"
  ],
  "maxResults": 100
}
```

# Actor output Schema

## `results` (type: `string`):

The dataset with one flat record per result. Append ?format=csv or ?format=xlsx to the URL for other formats.

## `summary` (type: `string`):

OUTPUT record in the key-value store: counts of results pushed and charged, requests, retries and duration.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "domains": [
        "allbirds.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("brightpath-data/shopify-store-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "domains": ["allbirds.com"] }

# Run the Actor and wait for it to finish
run = client.actor("brightpath-data/shopify-store-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "domains": [
    "allbirds.com"
  ]
}' |
apify call brightpath-data/shopify-store-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,brightpath-data/shopify-store-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/UiA8hbiJJC5GVdGc6/builds/t53s6Yf0WyDH0KNoi/openapi.json
