# Zapier Apps Scraper (`datascrapers/zapier-apps-scraper`) Actor

Scrape Zapier’s app directory by category, subcategory, sort (popularity, premium, beta, recently launched), or URL. Optionally enrich each app with triggers, actions, and integrations.

- **URL**: https://apify.com/datascrapers/zapier-apps-scraper.md
- **Developed by:** [Farhan Ali](https://apify.com/datascrapers) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.20 / 1,000 app scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

**Zapier Apps Scraper** creates a structured dataset of app records collected from zapier.com. Each dataset item represents one app in the Zapier directory and can include the name, slug, website, description, premium and beta flags, partner tier, categories, and logo, with optional triggers, actions, searches, paired integrations, alternatives, and Zap templates. Query the source using a category, subcategory, app slug, sort filter, or a direct Zapier URL, control the result limit with `maxItems`, and retrieve records through the Apify Dataset API or export them as JSON, CSV, Excel, or XML.

### Dataset at a glance

| Property | Value |
|---|---|
| Source | zapier.com |
| Record unit | One Zapier app |
| Input methods | `category`, `subcategory`, `appIds`, `sortBy`, `startUrls` |
| Main identifiers | `id`, `slug`, `url` |
| Delivery | Apify Dataset and API |
| Export formats | JSON, CSV, Excel, XML |
| Update model | Fresh records per Actor run |
| Pricing | $1.50 per 1,000 apps |

### Coverage and available records

The Actor returns apps from Zapier's public directory. Supported coverage includes:

- Main directory categories (for example `marketing`, `productivity`, `artificial-intelligence`, or `all`).
- Optional child subcategories (for example `ads-conversion`, `ai-agents`, `spreadsheets`).
- Direct app slugs via `appIds` (for example `google-sheets`, `slack`, `gmail`).
- Sort filters: `popularity` (default), `premium`, `beta`, and `recentlyLaunched`.
- Direct directory, category, or app-integration URLs via `startUrls`.
- Optional app-profile enrichment (`extractAppDetails`) for triggers, actions, searches, paired integrations, alternatives, and Zap templates.

Triggers, actions, searches, integrations, alternatives, Zap templates, and FAQs appear only when `extractAppDetails` is enabled. If no input is set, the Actor scrapes all apps in the directory.

### Data dictionary

| Field | Type | Nullable | Description | Example |
|---|---:|---|---|---|
| `id` | string | No | Zapier app identifier | `"0d71fb90-f233-4ce0-bdb1-3a2c887dfadf"` |
| `legacyId` | integer | Yes | Legacy numeric identifier | `1498` |
| `name` | string | No | App name | `"Google Sheets"` |
| `slug` | string | No | App slug | `"google-sheets"` |
| `url` | string | No | App page URL | `"https://zapier.com/apps/google-sheets"` |
| `integrationsUrl` | string | Yes | Integrations page URL | `"https://zapier.com/apps/google-sheets/integrations"` |
| `description` | string | Yes | Directory description | `"Create, edit, and share spreadsheets..."` |
| `websiteUrl` | string | Yes | Official product website | `"http://sheets.google.com/"` |
| `isBeta` | boolean | Yes | Beta flag | `false` |
| `isPremium` | boolean | Yes | Premium flag | `false` |
| `isUpcoming` | boolean | Yes | Upcoming flag | `false` |
| `partnerTier` | string | Yes | Zapier partner tier | `"PLATINUM"` |
| `primaryColor` | string | Yes | Brand primary color | `"00a256"` |
| `logoUrl` | string | Yes | App logo URL | `"https://zapier-images.imgix.net/..."` |
| `logoMediumUrl` | string | Yes | Medium logo URL | `"https://cdn.zapier.com/img/..."` |
| `logoSmallUrl` | string | Yes | Small logo URL | `"https://cdn.zapier.com/img/..."` |
| `categories` | array | Yes | Category objects (id, title, slug, role) | `[{"id": "49", "title": "Spreadsheets", "slug": "spreadsheets"}]` |
| `categoryTitles` | array | Yes | Category display names | `["Google", "Spreadsheets"]` |
| `sourceCategory` | string | Yes | Source category name | `null` |
| `sourceCategorySlug` | string | Yes | Source category slug | `null` |
| `sortBy` | string | Yes | Sort filter used | `null` |
| `hasAppDetails` | boolean | Yes | Whether details were fetched | `true` |
| `selectedApi` | string | Yes | Selected API version string | `"GoogleSheetsV2CLIAPI@2.17.0"` |
| `triggers` | array | Yes | Trigger operations, when details are enabled | `[{"type": "read", "label": "New Spreadsheet Row", "key": "new_row"}]` |
| `actions` | array | Yes | Action operations, when details are enabled | `[...]` |
| `searches` | array | Yes | Search operations, when details are enabled | `[...]` |
| `integrations` | array | Yes | Paired apps, when details are enabled | `[...]` |
| `alternatives` | array | Yes | Suggested alternative apps, when details are enabled | `[...]` |
| `zapTemplates` | array | Yes | Popular Zap templates, when details are enabled | `[...]` |
| `howItWorks` | string | Yes | Getting-started copy, when details are enabled | `"..."` |
| `faqs` | array | Yes | App FAQs, when details are enabled | `[...]` |

The most stable field for deduplication is `slug` (falling back to `url`).

### Example dataset record

```json
{
  "id": "0d71fb90-f233-4ce0-bdb1-3a2c887dfadf",
  "legacyId": 1498,
  "name": "Google Sheets",
  "slug": "google-sheets",
  "url": "https://zapier.com/apps/google-sheets",
  "integrationsUrl": "https://zapier.com/apps/google-sheets/integrations",
  "description": "Create, edit, and share spreadsheets wherever you are with Google Sheets, and get automated insights from your data.",
  "websiteUrl": "http://sheets.google.com/",
  "isBeta": false,
  "isPremium": false,
  "isUpcoming": false,
  "partnerTier": "PLATINUM",
  "primaryColor": "00a256",
  "logoUrl": "https://zapier-images.imgix.net/storage/services/8913a06feb7556d01285c052e4ad59d0.png",
  "categories": [
    { "id": "49", "title": "Spreadsheets", "slug": "spreadsheets", "role": "child", "type": "curated" }
  ],
  "categoryTitles": ["Google", "Spreadsheets"],
  "hasAppDetails": true,
  "triggers": [
    { "type": "read", "label": "New Spreadsheet Row (Team Drive)", "key": "new_row", "isHook": false }
  ]
}
```

This record was produced from the app directory with `extractAppDetails` enabled; trigger, action, and integration arrays are truncated here.

### Query and input reference

| Input | Type | Required | Default | Accepted values | Description |
|---|---:|---|---|---|---|
| `category` | string | No | `all` | `all`, `marketing`, `productivity`, `artificial-intelligence`, and others | Main directory category. |
| `subcategory` | string | No | none | Child category slugs | Optional child category. |
| `appIds` | array | No | none | App slugs (e.g. `google-sheets`, `slack`) | Direct app slugs to scrape. |
| `startUrls` | array | No | none | Zapier directory, category, or app URLs | Direct URLs; when set, `category` and `subcategory` are ignored. |
| `sortBy` | string | No | `popularity` | `popularity`, `premium`, `beta`, `recentlyLaunched` | Directory order/filter. |
| `extractAppDetails` | boolean | No | `false` | `true` / `false` | Fetch triggers, actions, integrations, and templates. |
| `detailConcurrency` | integer | No | 3 | 1–15 | Parallel app-profile enrichments. |
| `maxItems` | integer | No | 10 | `0` (unlimited) or a positive integer | Maximum apps to scrape. |
| `proxyConfiguration` | object | No | Apify Residential | Apify proxy settings | Proxy configuration. |

Minimal request:

```json
{
  "category": "marketing",
  "subcategory": "ads-conversion",
  "sortBy": "popularity",
  "maxItems": 5,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": ["RESIDENTIAL"]
  }
}
```

### Retrieve the data through the API

1. Start the Actor with a JSON input containing a category, subcategory, app slugs, or start URLs.
2. Wait for the run to finish, or use the synchronous run endpoint.
3. Retrieve items from the run's default dataset via the Dataset API.
4. Paginate the dataset or export it in the required format.

The Apify Console generates ready-to-run code for Python, JavaScript, and other languages in the Actor's API tab; see that tab for the current endpoint and authentication details. Never place a real API token in a URL or example.

### Data quality and record handling

- `triggers`, `actions`, `searches`, `integrations`, `alternatives`, `zapTemplates`, `howItWorks`, and `faqs` are present only when `extractAppDetails` is enabled.
- `isPremium`, `isBeta`, and `isUpcoming` reflect the directory flags at run time.
- Within a run, apps are deduplicated by `slug`. Across runs, key on `slug`.
- A single app that fails to fetch is skipped or written with listing-level fields; the Actor fails soft.
- Enable Apify Residential proxies (the default) and keep `detailConcurrency` modest to reduce blocks.

### Export and pipeline examples

| Destination | Recommended method | Typical use |
|---|---|---|
| PostgreSQL/Supabase | Dataset API or webhook consumer | Store app records keyed on `slug` |
| Google Sheets | Apify integration | Shareable integration catalog |
| S3/cloud storage | Scheduled export or integration | Periodic directory archive |

### Pricing and cost examples

Billing is pay-per-event. Each app written to the dataset is one `dataset-item` charge ($1.50 per 1,000); enabling app details adds a `listing-details` charge at $1.50 per 1,000. `apify-actor-start` is a small one-time per-run charge.

| Records | Estimated base cost |
|---:|---:|
| 1,000 apps | $1.50 |
| 10,000 apps | $15.00 |

Estimates assume the published pay-per-event model and do not include Apify platform usage; Apify subscription plan discounts may reduce the per-event rate. Enabling `extractAppDetails` increases cost per app.

### Limitations and responsible data use

- Only publicly accessible Zapier directory data is returned; operation lists require `extractAppDetails`.
- Premium/beta flags and partner tiers reflect the directory at run time and change over time.
- Extraction depends on Zapier's current availability and site structure, which can change without notice.
- No historical snapshots are stored unless the user archives dataset exports.
- Users are responsible for compliance with applicable terms of service and privacy obligations.

### Dataset questions

#### What does one dataset item represent?

One dataset item is a single app in the Zapier directory, identified by `slug`.

#### Which field should I use as a unique identifier?

Use `slug`; it is unique per app and stable across runs.

#### Are fields nullable or conditional?

Yes. `triggers`, `actions`, `searches`, `integrations`, `alternatives`, and `zapTemplates` are present only when `extractAppDetails` is enabled; `websiteUrl` and `description` may be absent for some apps.

#### Can I retrieve the records as CSV or JSON?

Yes. The dataset can be exported as JSON, CSV, Excel, or XML from the Apify Console or Dataset API.

#### How do I paginate large datasets?

Raise `maxItems` (set `0` for unlimited) to collect more records in a single run, then paginate via the Dataset API.

#### What counts as a billable result?

Each app record written to the dataset is one `dataset-item` charge; enabling `extractAppDetails` adds one `listing-details` charge per enriched app.

### Related datasets from Data Scrapers

- [AppSumo Scraper](https://apify.com/datascrapers/appsumo-scraper) — SaaS deal records for adjacent SaaS-market research.
- [Clutch.co Company Scraper](https://apify.com/datascrapers/clutch-scraper) — B2B company profiles for vendor and market analysis.
- [LinkedIn Company Scraper](https://apify.com/datascrapers/linkedin-company-scraper) — company records for firm enrichment.
- [Play Store App Reviews](https://apify.com/datascrapers/playstore-app-reviews) — app review records for app-market research.

### Data Scrapers support

Need an additional field, record type, or export workflow? Contact Data Scrapers at stardustspotlight@gmail.com. Include a sample source URL, required fields, expected record volume, and preferred delivery format.

# Actor input Schema

## `category` (type: `string`):

Zapier apps directory category to scrape (e.g. Artificial Intelligence, Marketing, Productivity). Default All apps scrapes the full directory. Ignored when Start URLs are provided. Combined with App slugs only when a specific category is selected.

## `subcategory` (type: `string`):

Optional child category under the main category (e.g. Ads & Conversion, Task Management, AI Agents). Leave empty to scrape the whole main category. Ignored when Start URLs are provided.

## `appIds` (type: `array`):

Zapier app slugs to scrape directly (e.g. google-sheets, slack, gmail). Primary input for AI agents — no URL required. Combined with Main category when a specific category is selected.

## `startUrls` (type: `array`):

Zapier directory or app URLs (e.g. https://zapier.com/apps, https://zapier.com/apps/categories/marketing, https://zapier.com/apps/google-sheets/integrations). When provided, Category and Subcategory are ignored.

## `sortBy` (type: `string`):

How to order directory results: Popularity (default), Premium apps only, Beta apps, or Recently launched. Applies to category listings and category Start URLs.

## `extractAppDetails` (type: `boolean`):

When enabled, fetch each app profile for triggers, actions, searches, paired integrations, alternatives, and Zap templates. Charges the listing-details event in addition to dataset-item. Off by default for faster, cheaper listing-only runs.

## `detailConcurrency` (type: `integer`):

How many app profiles to enrich in parallel when Extract app details is enabled (e.g. 3, 5). Higher is faster; keep it modest to avoid blocks.

## `maxItems` (type: `integer`):

Maximum number of apps to scrape (0 = unlimited)

## `proxyConfiguration` (type: `object`):

Proxy settings for anti-bot protection. Apify Residential proxy is recommended.

## Actor input object example

```json
{
  "category": "all",
  "subcategory": "",
  "appIds": [
    "google-sheets",
    "slack"
  ],
  "startUrls": [
    {
      "url": "https://zapier.com/apps/categories/marketing"
    }
  ],
  "sortBy": "popularity",
  "extractAppDetails": false,
  "detailConcurrency": 3,
  "maxItems": 10,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# Actor output Schema

## `results` (type: `string`):

Dataset of scraped Zapier apps

## `runStats` (type: `string`):

Record counts and run timestamps

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "appIds": [
        "google-sheets",
        "slack"
    ],
    "startUrls": [
        {
            "url": "https://zapier.com/apps/categories/marketing"
        }
    ],
    "detailConcurrency": 3,
    "maxItems": 10,
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("datascrapers/zapier-apps-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "appIds": [
        "google-sheets",
        "slack",
    ],
    "startUrls": [{ "url": "https://zapier.com/apps/categories/marketing" }],
    "detailConcurrency": 3,
    "maxItems": 10,
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("datascrapers/zapier-apps-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "appIds": [
    "google-sheets",
    "slack"
  ],
  "startUrls": [
    {
      "url": "https://zapier.com/apps/categories/marketing"
    }
  ],
  "detailConcurrency": 3,
  "maxItems": 10,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call datascrapers/zapier-apps-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,datascrapers/zapier-apps-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/k9RE33YotFbxwbexZ/builds/OfCPrmtgKSS072Suu/openapi.json
