# The Warehouse NZ Product Scraper (`acquistion-automation/thewarehouse-nz-scraper`) Actor

Scrapes product listings from The Warehouse New Zealand by search query or category path. Returns each product as a flat row with price, rating, availability, and full details.

- **URL**: https://apify.com/acquistion-automation/thewarehouse-nz-scraper.md
- **Developed by:** [Acquisition Automation Co.](https://apify.com/acquistion-automation) (community)
- **Categories:** E-commerce, Automation, Integrations
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $7.50 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

### The Warehouse NZ Product Scraper

**Scrape product listings from The Warehouse NZ by search term or category, up to a million per run.** Each product comes with its price, availability, rating, and full details. No login or API key. Export to CSV, JSON, Excel, or XML.

The Warehouse NZ's website holds thousands of products across home, electronics, toys, and clothing, but browsing manually or checking stock one page at a time is slow. This scraper reads the public search and category feeds directly, so you can collect structured product data for price monitoring, assortment analysis, or market research.
It returns every matched product in a consistent flat schema, ready for spreadsheets or databases.

| Who uses it | What they scrape The Warehouse NZ for |
|---|---|
| Pricing analysts | Track daily price changes on specific product categories. |
| E-commerce managers | Monitor competitor assortment and stock availability in New Zealand. |
| Market researchers | Understand product trends and pricing strategies in the NZ retail market. |
| Dropshippers | Build a product feed with current prices and stock status for their store. |

### What it does

This Actor collects product listings from The Warehouse NZ by search query or category path, and returns each one as a flat row with price, rating, and availability.

- 🔍 **Search by keyword:** enter any search term like 'kettle' or 'kids shoes' to collect matching products.
- 📂 **Filter by category:** provide a category path slug from a Warehouse URL to scrape a whole department.
- 🔢 **Sort control:** order results by relevance, price ascending or descending, newest, or top rated.
- 📊 **Structured output:** every product returns as a flat row with price, rating, availability, and product details.

Results export to CSV, JSON, Excel, or XML, or straight from the API.

### What you can do with The Warehouse NZ data

**💰 Monitor competitor pricing.**

A pricing analyst runs the scraper daily on key categories to track how The Warehouse adjusts prices on electronics and home goods.

**📦 Check stock availability.**

A dropshipper scrapes their product list weekly to confirm which items are in stock and remove discontinued lines from their feed.

**📈 Analyze product assortment.**

A market researcher collects the full toy category before Christmas to see which brands and price points dominate the shelves.

**🛒 Build a product catalogue.**

An e-commerce manager scrapes search results for 'outdoor furniture' to populate a comparison site with current NZ dollar prices.

### Why choose this scraper

|  | What you get |
|---|---|
| **No API key needed** | Access public product data without registration or rate limits. |
| **Flexible input** | Search by keyword or scrape entire category pages with a path slug. |
| **Sort on the fly** | Get results ordered by price, rating, or newest items first. |
| **Scalable collection** | Collect up to a million products in a single run. |

### How it compares

No other Store actor targets The Warehouse NZ the same way, so the honest comparison is with the alternatives teams actually weigh.

| | The Warehouse NZ Product Scraper | Build it in-house | By hand |
|---|---|---|---|
| Setup | Run it now, zero config | Days of engineering | None, but hours per pull |
| When The Warehouse NZ changes | Maintained for you | You fix it | You re-learn the page |
| Proxies, retries, anti-bot | Built in | Your problem | Browser only |
| Output | Fixed JSON schema, CSV/Excel export | Whatever you build | Copy-paste |
| Cost | Pay per result | Engineering time | Analyst hours |

### Configure the run

Drive the Actor from search queries and category paths, alone or together, and sorting runs before collection so only the most relevant products reach your dataset. The Input tab lists every parameter.

A first run with the defaults:

```json
{
  "query": "kettle",
  "maxItems": 10
}
```

A larger pull:

```json
{
  "query": "kettle",
  "maxItems": 200
}
```

### Pricing

Pay-per-result: **$0.0085 per result** collected. You pay only for the results written to your dataset.

| Results collected | Approximate cost |
|---|---|
| 100 results | $0.85 |
| 1,000 results | $8.50 |
| 10,000 results | $85.00 |

New Apify accounts start with $5 in free credit.

### Free users

Free-plan runs return up to 10 results as a preview. [Upgrade your Apify plan](https://console.apify.com/sign-up) to collect up to 1,000,000 results per run.

### Run it

1. [Create a free Apify account with $5 in credit](https://console.apify.com/sign-up).
2. Open the [The Warehouse NZ Product Scraper](https://apify.com/acquistion-automation/thewarehouse-nz-scraper).
3. Set your inputs and any filters, then click **Start**.
4. Export the results as CSV, Excel, JSON, or XML from the **Dataset** tab.

Run it programmatically through the [Apify API](https://docs.apify.com/api/v2) (`run-sync-get-dataset-items`) or the [ApifyClient](https://docs.apify.com/api/client/js) for JavaScript and Python.

### Use with AI agents (MCP)

Give an AI agent live access to The Warehouse NZ through the Model Context Protocol. Add the Actor to Claude, Cursor, or any MCP client:

```bash
claude mcp add --transport http apify "https://mcp.apify.com?tools=acquistion-automation/thewarehouse-nz-scraper"
```

Then prompt it in plain language to run the scraper and read back the results.

### Troubleshooting

**Why am I getting no results?**

Check your search term spelling or try a broader keyword. If using a category path, verify it matches a live URL on The Warehouse site. Some categories may have no products listed.

**The scraper stopped before reaching my max items limit.**

The scraper collects all products that match your query. If fewer products exist than your max items setting, it will finish early. This is expected behavior.

**Some product prices are missing or show as zero.**

The Warehouse website sometimes hides prices until you select a variant or size. The scraper captures the price as displayed on the search or category page. Check the product page directly if a price looks incorrect.

**The sort order does not match what I see on the website.**

The scraper sends the sort parameter to The Warehouse's API. Small differences can occur if the website applies additional client-side filtering. Try a different sort option or collect unsorted data and sort it yourself.

### FAQ

| Question | Answer |
|---|---|
| Can I scrape a specific category like 'toys' or 'clothing'? | Yes. Copy the category path slug from any Warehouse URL and paste it into the category input field. The scraper will collect all products listed in that department. |
| How many products can I collect in one run? | You can set the maximum up to 1,000,000 products. The actual number collected depends on how many match your search or category. |
| Does this scraper handle The Warehouse's weekly deals or clearance items? | It collects whatever products appear for your search or category, including items on sale or clearance. The output includes the current price and any special offer labels. |
| What output formats are supported? | You can export your dataset as CSV, JSON, Excel, or XML directly from Apify's storage. |
| Do I need an account or API key from The Warehouse? | No. This scraper reads the publicly accessible product feeds. You do not need a Warehouse account or any API registration. |
| Can I sort results by price or newest items? | Yes. Use the sort dropdown to order results by relevance, price low to high, price high to low, newest, or top rated before collection begins. |
| Is it legal to scrape The Warehouse website? | This Actor only accesses publicly available data. You are responsible for complying with The Warehouse's terms of service and local laws in your jurisdiction. |
| Can I run this scraper on a schedule? | Yes. Apify supports scheduled runs. You can set it to collect product data hourly, daily, or weekly to track changes over time. |

🆘 **Need help?** Open an issue in the Issues tab of this Actor with your run ID, your input, and what you expected.

⚠️ **Disclaimer.** This Actor is unofficial and is not affiliated with, endorsed by, or sponsored by The Warehouse Group Limited. It collects only publicly available data. You are responsible for using the collected data in compliance with the source's terms of service and applicable data-protection laws, including GDPR, CCPA, and PIPL. Do not use it to collect personal data unlawfully.

# Actor input Schema

## `query` (type: `string`):

Keyword to search The Warehouse (e.g. 'kettle', 'kids shoes').

## `maxItems` (type: `integer`):

Maximum number of products to collect per run.

## `category` (type: `string`):

Optional category path slug from a Warehouse URL.

## `sort` (type: `string`):

Sort by

## Actor input object example

```json
{
  "query": "kettle",
  "maxItems": 10,
  "sort": "relevance"
}
```

# Actor output Schema

## `results` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "query": "kettle",
    "maxItems": 10
};

// Run the Actor and wait for it to finish
const run = await client.actor("acquistion-automation/thewarehouse-nz-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "query": "kettle",
    "maxItems": 10,
}

# Run the Actor and wait for it to finish
run = client.actor("acquistion-automation/thewarehouse-nz-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "query": "kettle",
  "maxItems": 10
}' |
apify call acquistion-automation/thewarehouse-nz-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,acquistion-automation/thewarehouse-nz-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/tyLvVtr3IDhkhXeKD/builds/OdKWhnT4wQaPUfIcm/openapi.json
