# MedicalExpo Scraper - Medical Devices & Specs (`crawloop/medicalexpo-scraper`) Actor

Scrape MedicalExpo medical devices for OEMs and buyers. Get title, model, manufacturer, characteristics, specs, images and PDF catalogs. Keyword, listing, stand or PDP — MedicalExpo API alternative.

- **URL**: https://apify.com/crawloop/medicalexpo-scraper.md
- **Developed by:** [Andrej Kiva](https://apify.com/crawloop) (community)
- **Categories:** E-commerce, Lead generation, AI
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.80 / 1,000 product records

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## MedicalExpo Scraper — Medical Devices & Manufacturer Specs

> **Disclaimer:** Unofficial integration for publicly accessible sources. Trademarks belong to their respective owners. Provided for informational use only; users must comply with applicable platform terms and laws.

> **Crawloop B2B medical data** — MedicalExpo device catalogs plus company directories (Europages / WLW).

| MedicalExpo (device catalog) | Europages (EU directory) | WLW (DACH directory) |
| :--- | :--- | :--- |
| **MedicalExpo Scraper** ◄── you are here | [Europages Scraper](https://apify.com/crawloop/europages-scraper) | [WLW Scraper](https://apify.com/crawloop/wlw-scraper) |
| Medical devices, characteristics, PDF catalogs, manufacturers | EU companies, VAT, contacts | DE / AT / CH suppliers |

**MedicalExpo scraper** for Apify — a practical **MedicalExpo API alternative** that turns the medical device marketplace into structured JSON. Scrape **product title, model, manufacturer, clinical characteristics, specifications, images, PDF catalogs**, and **company websites** from keywords, category listings, manufacturer stands, or product URLs.

Built for **medical device competitive intelligence**, **hospital and distributor sourcing**, **OEM supplier discovery**, and **catalog monitoring**. Run from the Console, **Python**, **Node.js**, or **MCP**. Fast HTTP crawl via `curl_cffi` — parses VirtualExpo `__preloadData__` (no headless browser).

### Use cases

| Use case | What you get |
| :--- | :--- |
| **Device shortlists by type** | Keyword → medical-manufacturer listing → product rows (ultrasound, defibrillator, OR table, …) |
| **Characteristic comparison** | Application, Ergonomics, Technology and other feature rows for side-by-side research |
| **Manufacturer catalog pull** | All products on a stand URL with optional full PDP enrichment |
| **PDF catalog harvest** | Linked datasheet / brochure titles and viewer URLs |
| **Website enrichment** | External manufacturer website + off-platform product link when published |
| **Category deep-dive** | `/cat/` pages expand into child product-type listings |

### When to use this Actor

- You need a **MedicalExpo scraper** that returns dataset rows (not just company contacts)
- You have **keywords**, **listing URLs**, **manufacturer stands**, or **product PDPs**
- You want **characteristics, specs, images, and PDF catalogs** in one JSON export
- You prefer a **browser-free** crawl callable from Python, Node.js, cURL, or MCP

### When not to use this Actor

- **Guaranteed live prices / stock** — most listings are RFQ / price-on-request
- **Sending RFQs** through the portal contact form — this Actor is read-only extraction
- **Company-directory firmographics (VAT, phone)** — use Europages or WLW instead
- **Trade-show exhibitor booth lists (e.g. MEDICA)** — different job; this Actor is the year-round MedicalExpo catalog
- **Authenticated MySpace-only fields** — public pages only

### Key features

- **Keyword search** — resolves via MedicalExpo kwref sitemaps to listing URLs
- **Listing & category URLs** — paginated `medical-manufacturer` pages; `/cat/` expands to children
- **Manufacturer stands** — crawl all product cards on a company stand
- **Product detail enrichment** — `fetchDetails` parses `__preloadData__` for full specs
- **VirtualExpo portal switch** — optional `portal` for sister hosts (DirectIndustry, AeroExpo, …)
- **Streaming results** — dataset rows appear while the run is in progress
- **Deduped push** — unique by `portal` + `productId` within a run
- **Pagination controls** — `maxPages` + `maxItems` for predictable run size
- **Lightweight & resilient** — Chrome TLS fingerprinting, proxy session rotation on WAF challenges

### Input parameters

| Parameter | Type | Default | Description |
| :--- | :--- | :--- | :--- |
| `searchKeywords` | Array | `["ultrasound system"]` | Keywords → kwref listing URLs. |
| `startUrls` | Array | `[]` | Mixed product / manufacturer / listing / category URLs. |
| `listingUrls` | Array | `[]` | medical-manufacturer and/or `/cat/` URLs. |
| `productUrls` | Array | `[]` | Direct product detail URLs. |
| `manufacturerUrls` | Array | `[]` | Manufacturer stand URLs. |
| `portal` | String | `"medicalexpo"` | VirtualExpo host (`medicalexpo`, `directindustry`, …). |
| `fetchDetails` | Boolean | `true` | Open PDPs for specs, description, images, catalogs. |
| `maxItems` | Integer | `50` | Max dataset rows (`0` = unlimited within `maxPages`). |
| `maxPages` | Integer | `3` | Max listing pages per list URL. |
| `concurrency` | Integer | `3` | Parallel PDP workers (1–15). |
| `proxyConfiguration` | Object | residential FR | Apify Proxy — residential + country `FR` recommended. |

#### Example — keyword product crawl

```json
{
  "searchKeywords": ["ultrasound system"],
  "fetchDetails": true,
  "maxItems": 50,
  "maxPages": 3,
  "concurrency": 3,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": ["RESIDENTIAL"],
    "apifyProxyCountry": "FR"
  }
}
```

#### Example — manufacturer stand + product URLs

```json
{
  "manufacturerUrls": [
    { "url": "https://www.medicalexpo.com/prod/mindray-north-america-70629.html" }
  ],
  "productUrls": [
    { "url": "https://www.medicalexpo.com/prod/mindray-north-america/product-70629-2798311.html" }
  ],
  "fetchDetails": true,
  "maxItems": 100,
  "concurrency": 3
}
```

### Output

Each dataset item is one medical device product.

| Field | Description |
| :--- | :--- |
| `title` / `model` | Product label and model designation |
| `companyName` / `companyId` | Manufacturer stand name and id |
| `url` / `companyUrl` | Product PDP and manufacturer stand URLs |
| `companyWebsite` | External manufacturer website when published |
| `features` | Characteristic rows (Application, Ergonomics, …) |
| `specifications` | Numeric / range specs (`min` / `max` / `raw`) |
| `images` | Product image URLs |
| `catalogs` | Linked PDF catalog title, URL, pages, language |
| `description` | Full product description from the detail page |
| `category` / `breadcrumbs` | Listing category and navigation path |
| `enriched` | `true` when fields come from a product detail page |

Example (illustrative):

```json
{
  "recordType": "product",
  "portal": "medicalexpo",
  "productId": "2798311",
  "companyId": "70629",
  "title": "Liver steatosis ultrasound elastography system",
  "model": "Hepatus 6",
  "companyName": "Mindray North America",
  "url": "https://www.medicalexpo.com/prod/mindray-north-america/product-70629-2798311.html",
  "companyUrl": "https://www.medicalexpo.com/prod/mindray-north-america-70629.html",
  "description": "Integrates quantitative ViTE elastography …",
  "category": "Ultrasound elastography system",
  "features": [
    { "name": "Appplication", "value": "liver steatosis, liver fibrosis" },
    { "name": "Ergonomics", "value": "on casters, portable" }
  ],
  "images": ["https://img.medicalexpo.com/images_me/photo-g/70629-….jpg"],
  "enriched": true,
  "scrapedAt": "2026-08-06T12:00:00Z"
}
```

### Integration examples

#### Node.js

```javascript
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: process.env.APIFY_TOKEN });
const run = await client.actor('crawloop/medicalexpo-scraper').call({
  searchKeywords: ['ultrasound system'],
  fetchDetails: true,
  maxItems: 50,
  proxyConfiguration: {
    useApifyProxy: true,
    apifyProxyGroups: ['RESIDENTIAL'],
    apifyProxyCountry: 'FR',
  },
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items.slice(0, 5));
```

#### Python

```python
from apify_client import ApifyClient

client = ApifyClient(token)
run = client.actor("crawloop/medicalexpo-scraper").call(
    run_input={
        "searchKeywords": ["ultrasound system"],
        "fetchDetails": True,
        "maxItems": 50,
        "proxyConfiguration": {
            "useApifyProxy": True,
            "apifyProxyGroups": ["RESIDENTIAL"],
            "apifyProxyCountry": "FR",
        },
    }
)
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item.get("title"), item.get("model"), item.get("companyName"))
```

#### cURL

```bash
curl "https://api.apify.com/v2/acts/crawloop~medicalexpo-scraper/runs?token=$APIFY_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"searchKeywords":["ultrasound system"],"fetchDetails":true,"maxItems":50,"proxyConfiguration":{"useApifyProxy":true,"apifyProxyGroups":["RESIDENTIAL"],"apifyProxyCountry":"FR"}}'
```

### MCP and AI assistants

Use this Actor from AI tools via [Apify MCP](https://docs.apify.com/platform/integrations/mcp). Connect your Apify account, then call `crawloop/medicalexpo-scraper`.

Example prompts:

- "Run MedicalExpo Scraper for keyword ultrasound system, max 30, return title, model, companyName, features as JSON"
- "Scrape a MedicalExpo manufacturer stand URL and summarize imaging-related characteristics"
- "Chain MedicalExpo Scraper then Europages Scraper to map device catalogs to EU company contacts"

### Suite next step

For Europe-wide **company contacts / VAT / firmographics**, run [Europages Scraper](https://apify.com/crawloop/europages-scraper). For DACH-only suppliers, use [WLW Scraper](https://apify.com/crawloop/wlw-scraper).

### Related Actors

| Actor | Use for |
| :--- | :--- |
| **MedicalExpo Scraper** ◄── you are here | Medical device catalog, characteristics, PDF catalogs |
| [Europages Scraper](https://apify.com/crawloop/europages-scraper) | Europe-wide B2B directory, multi-locale, VAT & contacts |
| [WLW Scraper](https://apify.com/crawloop/wlw-scraper) | DACH (DE / AT / CH) B2B suppliers from Wer liefert was |

### FAQ

**Is this a MedicalExpo API?**\
No official public product API is required. This Actor is a **MedicalExpo scraper / API alternative** that returns structured dataset rows you can call from Python, Node.js, cURL, or MCP.

**How do I scrape MedicalExpo with Python or Node.js?**\
Use the Apify client examples above, or call the Actor from an AI assistant via Apify MCP. Export the default dataset as JSON, CSV, or Excel.

**Does keyword search need exact listing URLs?**\
No — keywords are matched against MedicalExpo kwref sitemaps (e.g. `ultrasound` → ultrasound listing). Prefer exact listing URLs when you already have them.

**Can I scrape DirectIndustry with the same Actor?**\
Set `portal` to `directindustry` (and sister portals). Page patterns are shared across VirtualExpo; MedicalExpo is the primary tested host for this Actor.

**Why are some rows missing phone/email or price?**\
This Actor targets **product catalog** fields. MedicalExpo is RFQ-oriented — public prices and manufacturer phones are usually not on product pages. Enrich firmographics with Europages or WLW.

# Actor input Schema

## `searchKeywords` (type: `array`):

Product-type keywords resolved via MedicalExpo kwref sitemaps to medical-manufacturer listing pages (e.g. "ultrasound system", "defibrillator", "operating table").

## `startUrls` (type: `array`):

Mix of product PDPs, manufacturer stands, medical-manufacturer listings, and/or category pages.

## `listingUrls` (type: `array`):

medical-manufacturer listing pages and/or /cat/ category pages (categories expand to child listings).

## `productUrls` (type: `array`):

Direct product detail URLs (/prod/{brand}/product-{companyId}-{productId}.html).

## `manufacturerUrls` (type: `array`):

Manufacturer stand URLs (/prod/{brand}-{companyId}.html) — expands to product cards on the stand.

## `portal` (type: `string`):

Which VirtualExpo marketplace host to use. MedicalExpo is the primary catalog; sister portals share the same page patterns.

## `fetchDetails` (type: `boolean`):

When true, open each product PDP and parse window.**preloadData** for full specs, description, images and PDF catalogs. When false, emit listing-card fields only.

## `maxItems` (type: `integer`):

Hard cap on dataset rows (0 = unlimited within maxPages).

## `maxPages` (type: `integer`):

Max pagination pages per medical-manufacturer listing URL.

## `concurrency` (type: `integer`):

Parallel PDP workers (1–15). Keep low if Cloudflare challenges appear.

## `proxyConfiguration` (type: `object`):

Apify Proxy settings. Residential recommended.

## Actor input object example

```json
{
  "searchKeywords": [
    "ultrasound system"
  ],
  "startUrls": [],
  "listingUrls": [],
  "productUrls": [],
  "manufacturerUrls": [],
  "portal": "medicalexpo",
  "fetchDetails": true,
  "maxItems": 50,
  "maxPages": 3,
  "concurrency": 3,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "FR"
  }
}
```

# Actor output Schema

## `results` (type: `string`):

Default dataset items (product records).

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {};

// Run the Actor and wait for it to finish
const run = await client.actor("crawloop/medicalexpo-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {}

# Run the Actor and wait for it to finish
run = client.actor("crawloop/medicalexpo-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{}' |
apify call crawloop/medicalexpo-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,crawloop/medicalexpo-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/lgPMspRSvh2cZT95F/builds/RTZVb4cjEQtcbzdru/openapi.json
