# Hornbach NL Scraper — DIY Products, Tools & Prices (`studio-amba/hornbach-nl-scraper`) Actor

Scrapes products from Hornbach.nl with prices, ratings, stock info and specs. Supports search queries and category URLs.

- **URL**: https://apify.com/studio-amba/hornbach-nl-scraper.md
- **Developed by:** [Studio Amba](https://apify.com/studio-amba) (community)
- **Categories:** E-commerce
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.20 / 1,000 result scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Hornbach NL Scraper -- Dutch DIY Megastore for Project Builders

Scrape products, prices, ratings, stock status, and SKU data from hornbach.nl -- the Dutch storefront of the Hornbach DIY megastore chain. Extracts data from Hornbach's Apollo/GraphQL state for fast, reliable results.

### What is Hornbach NL Scraper?

Hornbach is not your average neighborhood hardware store. It's a warehouse-format DIY chain with large stores designed for customers tackling serious renovation projects, new builds, and professional contracting work. Hornbach runs stores across the Netherlands alongside its German parent chain and other European storefronts, and hornbach.nl is the dedicated Dutch site with Dutch-language listings and euro pricing.

Hornbach's website uses a modern React frontend with Apollo GraphQL state management. This scraper extracts product data directly from the server-rendered Apollo state embedded in the HTML -- no JavaScript rendering needed on the extraction side, which makes it fast and reliable once the page is fetched.

- **Contractor price benchmarking** -- Compare Hornbach NL's pricing on building materials, lumber, and insulation against Praxis, GAMMA, and Karwei
- **Bulk material cost tracking** -- Monitor prices on high-volume items like wood planks, tiles, screws by the box, and bulk paint to optimize procurement timing
- **Power tool competitive analysis** -- Track how Hornbach NL prices power tools from Bosch, Makita, DeWalt, and Metabo against other Dutch DIY chains
- **Garden and outdoor project planning** -- Extract pricing data for garden sheds, terrace materials, fencing, and outdoor equipment for project budgeting
- **Dutch DIY market research** -- Use hornbach.nl pricing as a data point alongside Hornbach's German and other European sites

### What data does Hornbach NL Scraper extract?

- **Product name** -- Full Dutch product title
- **Price & discounts** -- Current price, original/strike-through price when discounted
- **Brand** -- Manufacturer name (Bosch, Bosch Professional, Makita, DeWalt, HORNBACH, etc.)
- **Stock availability** -- Online availability status
- **Ratings & reviews** -- Average rating and review count
- **Product image** -- Main product image URL
- **Product URL** -- Direct link on hornbach.nl
- **SKU / Article number** -- Hornbach's internal article ID for reference and reordering

### How to scrape Hornbach Netherlands data

| Field | Type | Description |
|-------|------|-------------|
| `searchQuery` | String | Search by keyword -- e.g., `"boormachine"`, `"schroeven"`, `"tuinmeubelen"`, `"hout"` |
| `categoryUrl` | String | Direct category URL -- e.g., `https://www.hornbach.nl/c/machines-gereedschap-werkplaats/elektrisch-gereedschap/S5512/` |
| `maxResults` | Integer | Maximum products to scrape (default: 100, set to 0 for unlimited) |
| `proxyConfiguration` | Object | Apify proxy recommended for reliability |

**Tips:**

- `categoryUrl` overrides `searchQuery` if both are provided
- The search URL pattern is `https://www.hornbach.nl/s/{query}/`
- Category URLs use the pattern `/c/{category-path}/S{id}/` (note: this differs from hornbach.de, which uses `/shop/{Category}/S{id}/artikelliste.html`)
- Set `maxResults` to 0 to scrape everything the query or category returns
- Apify proxy is enabled by default; hornbach.nl runs the same anti-bot challenge as hornbach.de, so a Bright Data fallback kicks in automatically when the residential proxy is blocked

### Output

```json
{
    "name": "BOSCH Professional Klopboormachine GSB 13 RE",
    "brand": "Bosch Professional",
    "price": 89.95,
    "currency": "EUR",
    "url": "https://www.hornbach.nl/p/bosch-professional-klopboormachine-gsb-13-re/6094347/",
    "scrapedAt": "2026-08-27T14:00:00.000Z",
    "imageUrl": "https://media.hornbach.nl/hb/packshot/as.46756464.jpg?dvid=8",
    "rating": 3.1,
    "reviewCount": 21,
    "sku": "6094347_ST",
    "inStock": true
}
```

A building materials example:

```json
{
    "name": "Vurenhout geschaafd 18 x 95 x 3000 mm",
    "brand": "HORNBACH",
    "price": 12.45,
    "currency": "EUR",
    "url": "https://www.hornbach.nl/p/vurenhout-geschaafd-18x95x3000/5432109/",
    "scrapedAt": "2026-08-27T14:00:00.000Z",
    "imageUrl": "https://media.hornbach.nl/hb/packshot/as.12345678.jpg",
    "rating": 4.5,
    "reviewCount": 43,
    "sku": "5432109",
    "inStock": true
}
```

### How much does it cost?

Hornbach NL Scraper uses direct HTTP requests with Apollo state extraction, falling back to a JS-rendering fetch only when the anti-bot wall requires it -- no persistent browser overhead.

| Scenario | Products | Est. cost |
|----------|----------|-----------|
| Quick search | 100 | ~$0.15 |
| Category deep dive | 500 | ~$0.50 |
| Department inventory | 1,000 | ~$0.90 |
| Full catalog section | 5,000 | ~$3.50 |

Proxy costs may apply depending on your Apify plan and proxy configuration.

### Can I integrate?

- **JSON, CSV, Excel** -- Direct dataset download
- **Google Sheets** -- Auto-sync via Apify integration
- **Webhooks** -- Trigger your pipeline on completion
- **API** -- Programmatic access (see below)
- **Zapier / Make** -- 5,000+ app integrations
- **Amazon S3, Google Cloud Storage** -- Cloud storage push

### Can I use it as an API?

**Python:**

```python
from apify_client import ApifyClient

client = ApifyClient("YOUR_API_TOKEN")

run = client.actor("studio-amba/hornbach-nl-scraper").call(run_input={
    "searchQuery": "boormachine",
    "maxResults": 200,
    "proxyConfiguration": {"useApifyProxy": True}
})

for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(f"{item['name']} - EUR {item['price']}")
```

**JavaScript:**

```javascript
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: 'YOUR_API_TOKEN' });

const run = await client.actor('studio-amba/hornbach-nl-scraper').call({
    searchQuery: 'boormachine',
    maxResults: 200,
    proxyConfiguration: { useApifyProxy: true },
});

const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);
```

### FAQ

**Does this work for Hornbach in other countries?**
This scraper targets hornbach.nl (Netherlands). Hornbach also runs hornbach.de (Germany) and other European storefronts, each with its own URL structure -- the Dutch category path (`/c/{path}/S{id}/`) is different from the German one (`/shop/{Category}/S{id}/artikelliste.html`). Use Studio AMBA's separate Hornbach DE scraper for the German site.

**Why does Hornbach matter for project builders?**
Hornbach's warehouse format is built for renovation and construction projects rather than small household fixes. Stores stock materials in bulk quantities and carry professional-grade tools alongside DIY basics.

**How does the Apollo state extraction work?**
Hornbach's website server-renders product listings using React with Apollo Client. The full Apollo state (including product data, prices, ratings) is embedded in the HTML as `window.__APOLLO_STATE__`. The scraper parses this JSON directly, which is faster and more reliable than scraping HTML elements.

**What if the Apollo state is not found?**
The scraper includes an HTML fallback parser that extracts product data from DOM elements (product tiles, price elements, brand labels). This covers cases where Hornbach changes their rendering approach.

**Why does this actor need Bright Data as a fallback?**
hornbach.nl runs an F5/Shape-style "Client Challenge" JavaScript wall in front of the site. Datacenter and even residential Apify proxy exits are frequently challenged and return no Apollo state. When that happens, the actor automatically retries the same page through Bright Data's Web Unlocker with JS rendering enabled, which solves the challenge.

**Can I track Hornbach NL's weekly promotions?**
Yes. Products with active promotions will show both `price` (current) and `originalPrice` (before discount). Schedule the actor to run weekly to track promotional patterns over time.

**How large is Hornbach NL's online catalog?**
Individual categories can run from a few hundred to several thousand products -- the "elektrisch gereedschap" (power tools) category alone lists over 3,000 items. Set `maxResults` based on the scope of your query.

### Limitations

- Targets hornbach.nl (Netherlands) only
- Data is extracted from listing-level Apollo state, so detailed specs and descriptions are not included (these would require visiting individual product pages)
- If Hornbach changes their Apollo state structure or switches rendering approaches, extraction may be temporarily affected
- Pagination stops when fewer than 20 products are found on a page (heuristic for detecting the last page)
- Category and search pages sit behind an anti-bot challenge; when the built-in Bright Data fallback is exhausted or unavailable, some runs may return fewer results than expected

### Other DIY & home improvement scrapers

Compare Hornbach NL with other major European DIY retailers:

- **[Hornbach Scraper](https://apify.com/studio-amba/hornbach-scraper)** -- Germany's Hornbach DIY megastore
- **[Praxis Scraper](https://apify.com/studio-amba/praxis-scraper)** -- Netherlands, Maxeda DIY Group
- **[GAMMA Scraper](https://apify.com/studio-amba/gamma-scraper)** -- Belgium, Intergamma group
- **[Brico Scraper](https://apify.com/studio-amba/brico-scraper)** -- Belgium's largest DIY chain
- **[Hubo Scraper](https://apify.com/studio-amba/hubo-scraper)** -- Belgian neighborhood DIY
- **[OBI Scraper](https://apify.com/studio-amba/obi-scraper)** -- Germany's #1 DIY chain by store count
- **[Bauhaus Scraper](https://apify.com/studio-amba/bauhaus-scraper)** -- German/EU professional hardware & DIY
- **[Leroy Merlin Scraper](https://apify.com/studio-amba/leroymerlin-scraper)** -- France's #1 DIY retailer
- **[Castorama Scraper](https://apify.com/studio-amba/castorama-scraper)** -- France, Kingfisher group

### Your feedback

Found a bug? Need additional data fields? Questions about Hornbach's data structure? Open an issue on the actor's page or reach out through Apify.

# Actor input Schema

## `searchQuery` (type: `string`):

Search term to look up on Hornbach.nl (e.g., 'boormachine', 'schroeven', 'tuinmeubelen')

## `categoryUrl` (type: `string`):

Direct URL to a Hornbach.nl category or search results page. Overrides searchQuery if both are provided.

## `maxResults` (type: `integer`):

Maximum number of products to scrape. Set to 0 for unlimited.

## `proxyConfiguration` (type: `object`):

Apify proxy settings. Use Apify residential proxies if you get blocked.

## `brightDataApiKey` (type: `string`):

Optional: your own Bright Data API key for the Web Unlocker zone. Used as a fallback when Hornbach's anti-bot challenge blocks residential proxies. If not set, the actor uses its built-in key.

## Actor input object example

```json
{
  "searchQuery": "boormachine",
  "categoryUrl": "https://www.hornbach.nl/c/machines-gereedschap-werkplaats/elektrisch-gereedschap/S5512/",
  "maxResults": 100,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "NL"
  }
}
```

# Actor output Schema

## `results` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "searchQuery": "boormachine",
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ],
        "apifyProxyCountry": "NL"
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("studio-amba/hornbach-nl-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "searchQuery": "boormachine",
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
        "apifyProxyCountry": "NL",
    },
}

# Run the Actor and wait for it to finish
run = client.actor("studio-amba/hornbach-nl-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "searchQuery": "boormachine",
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "NL"
  }
}' |
apify call studio-amba/hornbach-nl-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,studio-amba/hornbach-nl-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/oWkEiiPdkv1arpAnE/builds/TZDt5SDleBaSf1zKl/openapi.json
