# Sephora Product Scraper — Prices, Shades & Ratings (`crawloop/sephora-scraper`) Actor

Sephora US product scraper for search keywords, category URLs, and product URLs. Get brand, price, rating, review count, and shade images on one row per product. HTTP — no browser. Built for price monitoring and assortment research.

- **URL**: https://apify.com/crawloop/sephora-scraper.md
- **Developed by:** [Andrej Kiva](https://apify.com/crawloop) (community)
- **Categories:** E-commerce, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $4.50 / 1,000 products

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Sephora Product Scraper — Prices, Shades & Ratings

> **Disclaimer:** Unofficial tool — not affiliated with, sponsored by, or endorsed by Sephora USA, Inc. or its affiliates. Data is read from publicly accessible US search and category listings only. No login. You are responsible for complying with applicable law and the site’s terms. No warranty on accuracy or availability. Provided for informational and research use.

Search the US catalog by **keyword**, **category URL**, or **product URL** and save one row per product: **brand**, **name**, **price**, **rating**, **review count**, and **shade images** nested on that row. This **Sephora scraper** is an **API alternative** for beauty **price monitoring** and assortment research. Run it from **Python**, **Node.js**, or **MCP**; export JSON / CSV / dataset. **HTTP** only — no headless browser.

> **Crawloop Marketplace & E-commerce Suite** — Sephora US listing cards, then Amazon, Walmart, or Flipkart for the same product outside this catalog.

| **Sephora Product Scraper** ◄── *you are here* | [Amazon Search Scraper](https://apify.com/crawloop/amazon-search-scraper) | [Walmart Scraper](https://apify.com/crawloop/walmart-scraper) | [Flipkart Scraper](https://apify.com/crawloop/flipkart-scraper) | [AliExpress Search Scraper](https://apify.com/crawloop/aliexpress-search-scraper) |
| :---: | :---: | :---: | :---: | :---: |
| US keyword and category cards, price, rating, shade images | Amazon keyword SERP, badges, sponsored | Walmart ZIP-local prices | Flipkart INR price, MRP, Assured | AliExpress prices, sold count, Choice |

***

### When to use this Actor

- **US price checks** — listed price or a price range on a keyword or category
- **Assortment pulls** — brand, name, rating, review count, and hero image
- **Shade counts** — how many color SKUs the listing exposes, with swatch image URLs when the category document includes them
- **Scheduled refresh** — same keyword, compare `price` and `rating` over time

### When not to use this Actor

- **Review text, ingredients, or store stock** — those are not on the listing document this Actor reads
- **Markets outside the US storefront** — this card is the US catalog
- **Amazon, Walmart, Flipkart, or AliExpress** — use the sibling Actors in the table above

***

### Key Features

- **Keyword search** — `query` or batch `queries`
- **Category URLs** — shop paths and category ids
- **Product URLs** — matched back to the listing that contains that product id (one row)
- **Shades on the product row** — SKU id and images nested in `variants`. One saved product is one row
- **Pagination + caps** — up to **20 pages** per query, or stop at `maxItems`
- **HTTP** — Chrome TLS impersonation, 256 MB. US residential proxy

***

### How to scrape Sephora products

1. Set a keyword in `query`, or paste US search, shop, or product URLs in `startUrls`.
2. Cap volume with `maxItems` and `maxPages`.
3. Keep the default US residential proxy.
4. Run the Actor. Download the default dataset as JSON, CSV, or Excel — or call it from Python, Node.js, or MCP.

***

### Input Parameters

| Parameter | Type | Default | Description |
| :--- | :--- | :--- | :--- |
| `query` | String | — | Primary US keyword |
| `queries` | Array | — | Extra keywords (same page cap) |
| `startUrls` | Array | — | US search, shop/category, or product URLs |
| `maxPages` | Integer | `1` | Pages per query or URL (1–20) |
| `maxItems` | Integer | — | Product cap per keyword or URL |
| `startPage` | Integer | `1` | First page |
| `deduplicateIds` | Boolean | `true` | One row per `productId` in the run |
| `maxConcurrency` | Integer | `1` | Parallel queries (1–2) |
| `proxyConfiguration` | Object | residential, US | Apify Proxy. Datacenter IPs are blocked |

#### Input example

```json
{
  "query": "lipstick",
  "maxPages": 1,
  "maxItems": 40,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": ["RESIDENTIAL"],
    "apifyProxyCountry": "US"
  }
}
```

***

### Output

Each dataset item is one product. Shade SKUs sit in `variants`.

| Field | Description |
| :--- | :--- |
| `productId`, `skuId` | Product id and the default SKU |
| `brand`, `name` | Brand and product name |
| `price`, `priceMax`, `listPrice` | Low price, high price when the card shows a range, and the original price text |
| `currency` | `USD` when the card shows a dollar price |
| `rating`, `reviewCount` | Average stars and review count |
| `variantCount`, `variants` | Shade count, plus SKU id and images when the listing includes swatches |
| `isBestseller`, `isNew`, `isLimitedEdition`, `isSephoraExclusive` | Flags on the default SKU |
| `imageUrl`, `url` | Hero image and public product URL |
| `page`, `position`, `searchQuery` | Where the row was found |

#### Output example

```json
{
  "source": "sephora_us",
  "searchQuery": "lipstick",
  "page": 1,
  "position": 1,
  "productId": "P510799",
  "skuId": "2837425",
  "brand": "MAC Cosmetics",
  "name": "M·A·Cximal Silky Matte Lipstick",
  "price": 16.0,
  "priceMax": 25.0,
  "listPrice": "$16.00 - $25.00",
  "currency": "USD",
  "rating": 4.5926,
  "reviewCount": 761,
  "variantCount": 26,
  "variationType": "Color",
  "isBestseller": true,
  "url": "https://www.sephora.com/product/mac-lipstick-P510799?skuId=2837425",
  "variants": [
    {
      "skuId": "2837425",
      "imageUrl": "https://www.sephora.com/productimages/sku/s2837425-main-zoom.jpg",
      "swatchImageUrl": "https://www.sephora.com/productimages/sku/s2837425+sw.jpg"
    }
  ],
  "scrapedAt": "2026-09-29T12:00:00Z"
}
```

***

### Use Cases

| Use case | What you get | Why it helps |
| :--- | :--- | :--- |
| **Price monitoring** | `price`, `priceMax`, `listPrice` | Track a keyword or category without opening each product |
| **Assortment build** | Brand, name, image, URL | Seed a sheet from search or a shop URL |
| **Shade coverage** | `variantCount` and `variants` | See how many color SKUs a listing exposes |
| **Cross-catalog check** | Same keyword on Amazon or Walmart | Compare a US beauty card with another retailer |

***

### Integration examples

#### Node.js

```javascript
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: process.env.APIFY_TOKEN });
const run = await client.actor('crawloop/sephora-scraper').call({
  query: 'lipstick',
  maxPages: 1,
  maxItems: 40
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items.slice(0, 5));
```

#### Python

```python
from apify_client import ApifyClient

client = ApifyClient(token)
run = client.actor("crawloop/sephora-scraper").call(
    run_input={"query": "lipstick", "maxPages": 1, "maxItems": 40}
)
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item.get("productId"), item.get("brand"), item.get("name"), item.get("price"))
```

#### cURL

```bash
curl "https://api.apify.com/v2/acts/crawloop~sephora-scraper/runs?token=$APIFY_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"query":"lipstick","maxPages":1,"maxItems":40}'
```

### MCP and AI assistants

Use this Actor from AI tools via [Apify MCP](https://docs.apify.com/platform/integrations/mcp). Connect your Apify account, then call this Actor by Store ID `crawloop/sephora-scraper`.

Example prompts:

- "Run Sephora Product Scraper for lipstick, 1 page, return productId, brand, name, price, rating, and reviewCount"
- "Scrape the US lipstick category and list bestsellers with more than 20 shades"
- "Chain Sephora Product Scraper then Amazon Search Scraper for the same keyword and compare US beauty prices with the Amazon SERP"

### Suite next step

After a Sephora keyword pull, compare the same products on [Amazon Search Scraper](https://apify.com/crawloop/amazon-search-scraper). For store prices by ZIP, use [Walmart Scraper](https://apify.com/crawloop/walmart-scraper).

***

### FAQ

**Does this open product pages?**
No. Keyword and category rows come from the listing document. A product URL is resolved by searching for that product id on the same listing, not by fetching a separate product page.

**Are review texts or ingredients included?**
No. The row has the rating and the review count shown on the listing. Review bodies, ingredients, and per-store stock are not collected.

**What if the run saves nothing?**
An empty keyword or a blocked session finishes without rows and without a crash. US residential proxy is required; datacenter IPs are blocked.

**Can I schedule it?**
Yes. Save a task with the same keyword and run it on a schedule. Compare `price` across runs.

# Actor input Schema

## `query` (type: `string`):

Sephora US keyword, for example lipstick.

## `queries` (type: `array`):

Optional extra keywords. Each query uses the same page cap.

## `startUrls` (type: `array`):

US search, shop/category, or product URLs. Product URLs are matched back to the listing that contains that product id.

## `maxPages` (type: `integer`):

Listing pages per keyword or URL (1–20). Stops earlier when maxItems is reached.

## `maxItems` (type: `integer`):

Cap on products saved per keyword or URL. A product URL saves at most one row.

## `startPage` (type: `integer`):

First results page (1-based).

## `deduplicateIds` (type: `boolean`):

Save each productId once across the whole run.

## `maxConcurrency` (type: `integer`):

How many keywords or URLs to fetch in parallel (1–2).

## `proxyConfiguration` (type: `object`):

Apify Proxy. US residential is required. Datacenter IPs are blocked.

## Actor input object example

```json
{
  "query": "lipstick",
  "maxPages": 1,
  "maxItems": 40,
  "startPage": 1,
  "deduplicateIds": true,
  "maxConcurrency": 1,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "US"
  }
}
```

# Actor output Schema

## `results` (type: `string`):

Default dataset items.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "query": "lipstick",
    "maxItems": 40
};

// Run the Actor and wait for it to finish
const run = await client.actor("crawloop/sephora-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "query": "lipstick",
    "maxItems": 40,
}

# Run the Actor and wait for it to finish
run = client.actor("crawloop/sephora-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "query": "lipstick",
  "maxItems": 40
}' |
apify call crawloop/sephora-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,crawloop/sephora-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/PnsLG0PaIK4wuEop2/builds/2EgvIjohyH7xVNYNP/openapi.json
