ArchiExpo Scraper - Architecture & Design Products avatar

ArchiExpo Scraper - Architecture & Design Products

Pricing

from $1.39 / 1,000 product records

Go to Apify Store
ArchiExpo Scraper - Architecture & Design Products

ArchiExpo Scraper - Architecture & Design Products

Scrape ArchiExpo architecture and design products for buyers and sourcing teams. Get title, manufacturer, characteristics, specs, images and PDF catalogs. Keyword, listing, stand or PDP — ArchiExpo API alternative.

Pricing

from $1.39 / 1,000 product records

Rating

0.0

(0)

Developer

Andrej Kiva

Andrej Kiva

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

8 days ago

Last modified

Share

ArchiExpo Scraper — Architecture & Design Products

Disclaimer: Unofficial integration for publicly accessible sources. Trademarks belong to their respective owners. Provided for informational use only; users must comply with applicable platform terms and laws.

Crawloop B2B architecture & design data — ArchiExpo product catalogs plus company directories (Europages / WLW).

ArchiExpo (architecture catalog)Europages (EU directory)WLW (DACH directory)
ArchiExpo Scraper ◄── you are hereEuropages ScraperWLW Scraper
Furniture, lighting, kitchen/bath, flooring, building materials, specs, PDF catalogsEU companies, VAT, contactsDE / AT / CH suppliers

ArchiExpo scraper for Apify — a practical ArchiExpo API alternative that turns the architecture and design marketplace into structured JSON. Scrape product title, model, manufacturer, characteristics, specifications, images, PDF catalogs, and company websites from keywords, category listings, manufacturer stands, or product URLs.

Built for furniture and lighting sourcing, kitchen / bath & flooring shortlists, building-materials competitive intelligence, and catalog monitoring. Run from the Console, Python, Node.js, or MCP. Fast HTTP crawl via curl_cffi — parses VirtualExpo __preloadData__ (no headless browser).

Use cases

Use caseWhat you get
Product shortlists by typeKeyword → architecture-design-manufacturer listing → product rows (sofa, lighting, kitchen sink, flooring, …)
Characteristic comparisonApplications, Material, Other characteristics and feature rows for side-by-side research
Manufacturer catalog pullAll products on a stand URL with optional full PDP enrichment
PDF catalog harvestLinked datasheet / brochure titles and viewer URLs
Website enrichmentExternal manufacturer website + off-platform product link when published
Category deep-dive/cat/ pages expand into child product-type listings

When to use this Actor

  • You need an ArchiExpo scraper that returns dataset rows (not just company contacts)
  • You have keywords, listing URLs, manufacturer stands, or product PDPs
  • You want characteristics, specs, images, and PDF catalogs in one JSON export
  • You prefer a browser-free crawl callable from Python, Node.js, cURL, or MCP

When not to use this Actor

  • Guaranteed live prices / stock — most listings are RFQ / price-on-request
  • Sending RFQs through the portal contact form — this Actor is read-only extraction
  • Company-directory firmographics (VAT, phone) — use Europages or WLW instead
  • Trade-show booth floor plans — different job; this Actor is the year-round ArchiExpo OEM / product catalog
  • Authenticated MySpace-only fields — public pages only

Key features

  • Keyword search — resolves via ArchiExpo kwref sitemaps to listing URLs
  • Listing & category URLs — paginated architecture-design-manufacturer pages; /cat/ expands to children
  • Manufacturer stands — crawl all product cards on a company stand
  • Product detail enrichmentfetchDetails parses __preloadData__ for full specs
  • VirtualExpo portal switch — optional portal for sister hosts (DirectIndustry, MedicalExpo, …)
  • Streaming results — dataset rows appear while the run is in progress
  • Deduped push — unique by portal + productId within a run
  • Pagination controlsmaxPages + maxItems for predictable run size
  • Lightweight & resilient — Chrome TLS fingerprinting, proxy session rotation on WAF challenges

Input parameters

ParameterTypeDefaultDescription
searchKeywordsArray["sofa"]Keywords → kwref listing URLs.
startUrlsArray[]Mixed product / manufacturer / listing / category URLs.
listingUrlsArray[]architecture-design-manufacturer and/or /cat/ URLs.
productUrlsArray[]Direct product detail URLs.
manufacturerUrlsArray[]Manufacturer stand URLs.
portalString"archiexpo"VirtualExpo host (archiexpo, directindustry, …).
fetchDetailsBooleantrueOpen PDPs for specs, description, images, catalogs.
maxItemsInteger50Max dataset rows (0 = unlimited within maxPages).
maxPagesInteger3Max listing pages per list URL.
concurrencyInteger3Parallel PDP workers (1–15).
proxyConfigurationObjectresidential FRApify Proxy — residential + country FR recommended.

Example — keyword product crawl

{
"searchKeywords": ["sofa"],
"fetchDetails": true,
"maxItems": 50,
"maxPages": 3,
"concurrency": 3,
"proxyConfiguration": {
"useApifyProxy": true,
"apifyProxyGroups": ["RESIDENTIAL"],
"apifyProxyCountry": "FR"
}
}

Example — manufacturer stand + product URLs

{
"manufacturerUrls": [
{ "url": "https://www.archiexpo.com/prod/alias-11279.html" }
],
"productUrls": [
{ "url": "https://www.archiexpo.com/prod/alias/product-11279-1810696.html" }
],
"fetchDetails": true,
"maxItems": 100,
"concurrency": 3
}

Output

Each dataset item is one architecture / design product.

FieldDescription
title / modelProduct label and model designation
companyName / companyIdManufacturer stand name and id
url / companyUrlProduct PDP and manufacturer stand URLs
companyWebsiteExternal manufacturer website when published
featuresCharacteristic rows (Applications, Material, …)
specificationsNumeric / range specs (min / max / raw)
imagesProduct image URLs
catalogsLinked PDF catalog title, URL, pages, language
descriptionFull product description from the detail page
category / breadcrumbsListing category and navigation path
enrichedtrue when fields come from a product detail page

Example (illustrative):

{
"recordType": "product",
"portal": "archiexpo",
"productId": "1810696",
"companyId": "11279",
"title": "Sofa",
"model": "ELEVEN PRIVACY 1 / 866",
"companyName": "ALIAS",
"url": "https://www.archiexpo.com/prod/alias/product-11279-1810696.html",
"companyUrl": "https://www.archiexpo.com/prod/alias-11279.html",
"companyWebsite": "http://alias.design",
"category": "Sofa",
"features": [
{ "name": "Capacity", "value": "2-person" },
{ "name": "Style", "value": "contemporary" }
],
"specifications": [
{ "name": "Seat height", "nominal": "44 cm (17.3 in)" }
],
"images": ["https://img.archiexpo.com/images_ae/photo-g/….jpg"],
"enriched": true,
"scrapedAt": "2026-08-11T12:00:00Z"
}

Integration examples

Node.js

import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: process.env.APIFY_TOKEN });
const run = await client.actor('crawloop/archiexpo-scraper').call({
searchKeywords: ['sofa'],
fetchDetails: true,
maxItems: 50,
proxyConfiguration: {
useApifyProxy: true,
apifyProxyGroups: ['RESIDENTIAL'],
apifyProxyCountry: 'FR',
},
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items.slice(0, 5));

Python

from apify_client import ApifyClient
client = ApifyClient(token)
run = client.actor("crawloop/archiexpo-scraper").call(
run_input={
"searchKeywords": ["sofa"],
"fetchDetails": True,
"maxItems": 50,
"proxyConfiguration": {
"useApifyProxy": True,
"apifyProxyGroups": ["RESIDENTIAL"],
"apifyProxyCountry": "FR",
},
}
)
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item.get("title"), item.get("model"), item.get("companyName"))

cURL

curl "https://api.apify.com/v2/acts/crawloop~archiexpo-scraper/runs?token=$APIFY_TOKEN" \
-H "Content-Type: application/json" \
-d '{"searchKeywords":["sofa"],"fetchDetails":true,"maxItems":50,"proxyConfiguration":{"useApifyProxy":true,"apifyProxyGroups":["RESIDENTIAL"],"apifyProxyCountry":"FR"}}'

MCP and AI assistants

Use this Actor from AI tools via Apify MCP. Connect your Apify account, then call crawloop/archiexpo-scraper.

Example prompts:

  • "Run ArchiExpo Scraper for keyword sofa, max 30, return title, companyName, features, specifications as JSON"
  • "Scrape ArchiExpo lighting listings and summarize manufacturers and material specs"
  • "Chain ArchiExpo Scraper then Europages Scraper to map architecture catalogs to EU company contacts"

Suite next step

For Europe-wide company contacts / VAT / firmographics, run Europages Scraper. For DACH-only suppliers, use WLW Scraper.

ActorUse for
ArchiExpo Scraper ◄── you are hereArchitecture & design catalog, characteristics, PDF catalogs
Europages ScraperEurope-wide B2B directory, multi-locale, VAT & contacts
WLW ScraperDACH (DE / AT / CH) B2B suppliers from Wer liefert was

FAQ

Is this an ArchiExpo API?
No official public product API is required. This Actor is an ArchiExpo scraper / API alternative that returns structured dataset rows you can call from Python, Node.js, cURL, or MCP.

How do I scrape ArchiExpo with Python or Node.js?
Use the Apify client examples above, or call the Actor from an AI assistant via Apify MCP. Export the default dataset as JSON, CSV, or Excel.

Does keyword search need exact listing URLs?
No — keywords are matched against ArchiExpo kwref sitemaps (e.g. sofa → sofa listing). Prefer exact listing URLs when you already have them.

Can I scrape DirectIndustry or MedicalExpo with the same Actor?
Set portal to directindustry or medicalexpo (and other sister portals). Page patterns are shared across VirtualExpo; ArchiExpo is the primary tested host for this Actor.

Why are some rows missing phone/email or price?
This Actor targets product catalog fields. ArchiExpo is RFQ-oriented — public prices and manufacturer phones are usually not on product pages. Enrich firmographics with Europages or WLW.