ArchiExpo Scraper - Architecture & Design Products
Pricing
from $1.39 / 1,000 product records
ArchiExpo Scraper - Architecture & Design Products
Scrape ArchiExpo architecture and design products for buyers and sourcing teams. Get title, manufacturer, characteristics, specs, images and PDF catalogs. Keyword, listing, stand or PDP — ArchiExpo API alternative.
Pricing
from $1.39 / 1,000 product records
Rating
0.0
(0)
Developer
Andrej Kiva
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
8 days ago
Last modified
Categories
Share
ArchiExpo Scraper — Architecture & Design Products
Disclaimer: Unofficial integration for publicly accessible sources. Trademarks belong to their respective owners. Provided for informational use only; users must comply with applicable platform terms and laws.
Crawloop B2B architecture & design data — ArchiExpo product catalogs plus company directories (Europages / WLW).
| ArchiExpo (architecture catalog) | Europages (EU directory) | WLW (DACH directory) |
|---|---|---|
| ArchiExpo Scraper ◄── you are here | Europages Scraper | WLW Scraper |
| Furniture, lighting, kitchen/bath, flooring, building materials, specs, PDF catalogs | EU companies, VAT, contacts | DE / AT / CH suppliers |
ArchiExpo scraper for Apify — a practical ArchiExpo API alternative that turns the architecture and design marketplace into structured JSON. Scrape product title, model, manufacturer, characteristics, specifications, images, PDF catalogs, and company websites from keywords, category listings, manufacturer stands, or product URLs.
Built for furniture and lighting sourcing, kitchen / bath & flooring shortlists, building-materials competitive intelligence, and catalog monitoring. Run from the Console, Python, Node.js, or MCP. Fast HTTP crawl via curl_cffi — parses VirtualExpo __preloadData__ (no headless browser).
Use cases
| Use case | What you get |
|---|---|
| Product shortlists by type | Keyword → architecture-design-manufacturer listing → product rows (sofa, lighting, kitchen sink, flooring, …) |
| Characteristic comparison | Applications, Material, Other characteristics and feature rows for side-by-side research |
| Manufacturer catalog pull | All products on a stand URL with optional full PDP enrichment |
| PDF catalog harvest | Linked datasheet / brochure titles and viewer URLs |
| Website enrichment | External manufacturer website + off-platform product link when published |
| Category deep-dive | /cat/ pages expand into child product-type listings |
When to use this Actor
- You need an ArchiExpo scraper that returns dataset rows (not just company contacts)
- You have keywords, listing URLs, manufacturer stands, or product PDPs
- You want characteristics, specs, images, and PDF catalogs in one JSON export
- You prefer a browser-free crawl callable from Python, Node.js, cURL, or MCP
When not to use this Actor
- Guaranteed live prices / stock — most listings are RFQ / price-on-request
- Sending RFQs through the portal contact form — this Actor is read-only extraction
- Company-directory firmographics (VAT, phone) — use Europages or WLW instead
- Trade-show booth floor plans — different job; this Actor is the year-round ArchiExpo OEM / product catalog
- Authenticated MySpace-only fields — public pages only
Key features
- Keyword search — resolves via ArchiExpo kwref sitemaps to listing URLs
- Listing & category URLs — paginated
architecture-design-manufacturerpages;/cat/expands to children - Manufacturer stands — crawl all product cards on a company stand
- Product detail enrichment —
fetchDetailsparses__preloadData__for full specs - VirtualExpo portal switch — optional
portalfor sister hosts (DirectIndustry, MedicalExpo, …) - Streaming results — dataset rows appear while the run is in progress
- Deduped push — unique by
portal+productIdwithin a run - Pagination controls —
maxPages+maxItemsfor predictable run size - Lightweight & resilient — Chrome TLS fingerprinting, proxy session rotation on WAF challenges
Input parameters
| Parameter | Type | Default | Description |
|---|---|---|---|
searchKeywords | Array | ["sofa"] | Keywords → kwref listing URLs. |
startUrls | Array | [] | Mixed product / manufacturer / listing / category URLs. |
listingUrls | Array | [] | architecture-design-manufacturer and/or /cat/ URLs. |
productUrls | Array | [] | Direct product detail URLs. |
manufacturerUrls | Array | [] | Manufacturer stand URLs. |
portal | String | "archiexpo" | VirtualExpo host (archiexpo, directindustry, …). |
fetchDetails | Boolean | true | Open PDPs for specs, description, images, catalogs. |
maxItems | Integer | 50 | Max dataset rows (0 = unlimited within maxPages). |
maxPages | Integer | 3 | Max listing pages per list URL. |
concurrency | Integer | 3 | Parallel PDP workers (1–15). |
proxyConfiguration | Object | residential FR | Apify Proxy — residential + country FR recommended. |
Example — keyword product crawl
{"searchKeywords": ["sofa"],"fetchDetails": true,"maxItems": 50,"maxPages": 3,"concurrency": 3,"proxyConfiguration": {"useApifyProxy": true,"apifyProxyGroups": ["RESIDENTIAL"],"apifyProxyCountry": "FR"}}
Example — manufacturer stand + product URLs
{"manufacturerUrls": [{ "url": "https://www.archiexpo.com/prod/alias-11279.html" }],"productUrls": [{ "url": "https://www.archiexpo.com/prod/alias/product-11279-1810696.html" }],"fetchDetails": true,"maxItems": 100,"concurrency": 3}
Output
Each dataset item is one architecture / design product.
| Field | Description |
|---|---|
title / model | Product label and model designation |
companyName / companyId | Manufacturer stand name and id |
url / companyUrl | Product PDP and manufacturer stand URLs |
companyWebsite | External manufacturer website when published |
features | Characteristic rows (Applications, Material, …) |
specifications | Numeric / range specs (min / max / raw) |
images | Product image URLs |
catalogs | Linked PDF catalog title, URL, pages, language |
description | Full product description from the detail page |
category / breadcrumbs | Listing category and navigation path |
enriched | true when fields come from a product detail page |
Example (illustrative):
{"recordType": "product","portal": "archiexpo","productId": "1810696","companyId": "11279","title": "Sofa","model": "ELEVEN PRIVACY 1 / 866","companyName": "ALIAS","url": "https://www.archiexpo.com/prod/alias/product-11279-1810696.html","companyUrl": "https://www.archiexpo.com/prod/alias-11279.html","companyWebsite": "http://alias.design","category": "Sofa","features": [{ "name": "Capacity", "value": "2-person" },{ "name": "Style", "value": "contemporary" }],"specifications": [{ "name": "Seat height", "nominal": "44 cm (17.3 in)" }],"images": ["https://img.archiexpo.com/images_ae/photo-g/….jpg"],"enriched": true,"scrapedAt": "2026-08-11T12:00:00Z"}
Integration examples
Node.js
import { ApifyClient } from 'apify-client';const client = new ApifyClient({ token: process.env.APIFY_TOKEN });const run = await client.actor('crawloop/archiexpo-scraper').call({searchKeywords: ['sofa'],fetchDetails: true,maxItems: 50,proxyConfiguration: {useApifyProxy: true,apifyProxyGroups: ['RESIDENTIAL'],apifyProxyCountry: 'FR',},});const { items } = await client.dataset(run.defaultDatasetId).listItems();console.log(items.slice(0, 5));
Python
from apify_client import ApifyClientclient = ApifyClient(token)run = client.actor("crawloop/archiexpo-scraper").call(run_input={"searchKeywords": ["sofa"],"fetchDetails": True,"maxItems": 50,"proxyConfiguration": {"useApifyProxy": True,"apifyProxyGroups": ["RESIDENTIAL"],"apifyProxyCountry": "FR",},})for item in client.dataset(run["defaultDatasetId"]).iterate_items():print(item.get("title"), item.get("model"), item.get("companyName"))
cURL
curl "https://api.apify.com/v2/acts/crawloop~archiexpo-scraper/runs?token=$APIFY_TOKEN" \-H "Content-Type: application/json" \-d '{"searchKeywords":["sofa"],"fetchDetails":true,"maxItems":50,"proxyConfiguration":{"useApifyProxy":true,"apifyProxyGroups":["RESIDENTIAL"],"apifyProxyCountry":"FR"}}'
MCP and AI assistants
Use this Actor from AI tools via Apify MCP. Connect your Apify account, then call crawloop/archiexpo-scraper.
Example prompts:
- "Run ArchiExpo Scraper for keyword sofa, max 30, return title, companyName, features, specifications as JSON"
- "Scrape ArchiExpo lighting listings and summarize manufacturers and material specs"
- "Chain ArchiExpo Scraper then Europages Scraper to map architecture catalogs to EU company contacts"
Suite next step
For Europe-wide company contacts / VAT / firmographics, run Europages Scraper. For DACH-only suppliers, use WLW Scraper.
Related Actors
| Actor | Use for |
|---|---|
| ArchiExpo Scraper ◄── you are here | Architecture & design catalog, characteristics, PDF catalogs |
| Europages Scraper | Europe-wide B2B directory, multi-locale, VAT & contacts |
| WLW Scraper | DACH (DE / AT / CH) B2B suppliers from Wer liefert was |
FAQ
Is this an ArchiExpo API?
No official public product API is required. This Actor is an ArchiExpo scraper / API alternative that returns structured dataset rows you can call from Python, Node.js, cURL, or MCP.
How do I scrape ArchiExpo with Python or Node.js?
Use the Apify client examples above, or call the Actor from an AI assistant via Apify MCP. Export the default dataset as JSON, CSV, or Excel.
Does keyword search need exact listing URLs?
No — keywords are matched against ArchiExpo kwref sitemaps (e.g. sofa → sofa listing). Prefer exact listing URLs when you already have them.
Can I scrape DirectIndustry or MedicalExpo with the same Actor?
Set portal to directindustry or medicalexpo (and other sister portals). Page patterns are shared across VirtualExpo; ArchiExpo is the primary tested host for this Actor.
Why are some rows missing phone/email or price?
This Actor targets product catalog fields. ArchiExpo is RFQ-oriented — public prices and manufacturer phones are usually not on product pages. Enrich firmographics with Europages or WLW.