# Clutch Scraper — Agency Profiles, Ratings & Reviews (`crawloop/clutch-scraper`) Actor

Scrape Clutch.co agency directories into JSON: ratings, reviews, hourly rates, min project size, team size, services, location and websites. Built for agency BD and procurement shortlists — no email enrichment. A Clutch API alternative for Python, Node.js, and MCP.

- **URL**: https://apify.com/crawloop/clutch-scraper.md
- **Developed by:** [Andrej Kiva](https://apify.com/crawloop) (community)
- **Categories:** Lead generation, AI
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.99 / 1,000 clutch companies

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Clutch Scraper — Agency Profiles, Ratings & Reviews

> **Disclaimer:** Unofficial integration for publicly accessible sources. Trademarks belong to their respective owners. Provided for informational use only; users must comply with applicable platform terms and laws.

> **Crawloop vendor-intel suite** — agency directories, public reviews, and employer signals for shortlists.

| Clutch (agencies) | Trustpilot (public reviews) | Glassdoor (employer) | ThomasNet (industrial) |
| :--- | :--- | :--- | :--- |
| **Clutch Scraper** ◄── you are here | [Trustpilot Scraper](https://apify.com/crawloop/trustpilot-scraper) | [Glassdoor Scraper](https://apify.com/crawloop/glassdoor-scraper) | [ThomasNet Scraper](https://apify.com/crawloop/thomasnet-scraper) |

**Clutch Scraper** for Apify — scrape **Clutch.co agency directories** into clean JSON without a public Clutch API. Extract **company name, rating, review count, hourly rate, min project size, team size, location, service mix, verification, sponsored vs organic rank, website, and optional client reviews**.

Built for **agency business development** (competitor maps, ranking, review-mined buyer context) and **procurement shortlists** (filters on rating, reviews, verified, budget bands). This Actor does **not** hunt emails on company websites. Run from the Console or with **Python**, **Node.js**, **cURL**, or **MCP** / AI assistants. A practical **Clutch API alternative**.

### When to use this Actor

- You need a **Clutch scraper** for a category or location directory (`/web-developers`, `/us/agencies/digital-marketing`)
- You want **organic rank vs sponsored** cards for competitor mapping
- You are building a **procurement shortlist** (min rating, min reviews, verified only)
- You need **client reviews** with project size, rating breakdown, and reviewer industry/location — not contact emails

### When not to use this Actor

- **Email enrichment / ValidatedMails / website crawling for inboxes** — out of scope on purpose
- **Posting projects or contacting firms through Clutch** — read-only extraction
- **Datacenter-only runs** — Cloudflare often blocks them; use Apify US Residential
- **The Manifest or other sister directories** — this Actor is Clutch.co only

### Key features

- **Clutch API alternative** — structured dataset instead of manual browsing
- **Directory + profile URLs** — category, country (`/us/...`), city slug, search, or `/profile/{slug}`
- **Procurement filters** — `minRating`, `minReviews`, `verifiedOnly`, `excludeSponsored`
- **Sponsored flag + listing position** — see paid vs earned rank on the page
- **Decoded website** — unwraps Clutch redirect links and strips referral UTMs
- **Optional reviews** — project size, schedule/quality/cost scores, reviewer firmographics
- **Past the first page** — `?page=N` with de-dupe by Clutch provider id (Clutch clamps overflowing pages)
- **Monitor** — scheduled NEW/UPDATED vs last run (rating, review count, rates, verified, organic rank). Skip UNCHANGED.
- **HTTP first, browser fallback** — curl\_cffi, then Camoufox if Cloudflare challenges

### Use cases

| Buyer | Job to be done |
| :--- | :--- |
| **Agency BD** | Map rivals in a category/geo: rates, review volume, service mix, organic position |
| **Procurement** | Shortlist vendors with rating/review floors, verified badge, min project size |
| **Market research** | Benchmark hourly bands and project sizes across a Clutch directory |
| **Due diligence** | Pull reviews (budget, timeline, quality) before an RFP — then run Trustpilot on the same domain |

### Input parameters

| Parameter | Type | Default | Description |
| :--- | :--- | :--- | :--- |
| `startUrls` | Array | Clutch web-developers | Directory, location, search, or profile URLs. |
| `category` | String | | Slug such as `web-developers` if you do not paste a URL. |
| `location` | String | | `us` / `uk` (prefix) or `new-york` (suffix on the category). |
| `fetchDetails` | Boolean | `false` | Open profiles for address, year founded, locations, social. |
| `includeReviews` | Boolean | `false` | Nest client reviews on each company row. |
| `maxReviewsPerCompany` | Integer | `50` | Review cap per profile (`0` = all paginated reviews). |
| `minRating` / `minReviews` | Number | | Procurement floors. |
| `verifiedOnly` | Boolean | `false` | Keep Clutch-verified companies only. |
| `excludeSponsored` | Boolean | `false` | Drop sponsor cards (organic rank). |
| `maxItems` | Integer | `50` | Max company rows (`0` = unlimited within `maxPages`). |
| `maxPages` | Integer | `5` | Max listing pages per start URL. |
| `incrementalMode` | Boolean | `false` | Emit NEW/UPDATED vs KV baseline; skip UNCHANGED. |
| `monitorBaselineOnly` | Boolean | `false` | Seed fingerprints on the first schedule without billing. |
| `proxyConfiguration` | Object | US residential | **Residential recommended.** |

#### Example — NYC web developers, organic shortlist

```json
{
  "startUrls": [{ "url": "https://clutch.co/web-developers/new-york" }],
  "maxItems": 50,
  "maxPages": 3,
  "excludeSponsored": true,
  "minRating": 4.5,
  "minReviews": 10,
  "fetchDetails": false,
  "includeReviews": false,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": ["RESIDENTIAL"],
    "apifyProxyCountry": "US"
  }
}
```

#### Example — profile + reviews (no emails)

```json
{
  "startUrls": [{ "url": "https://clutch.co/profile/lounge-lizard" }],
  "fetchDetails": true,
  "includeReviews": true,
  "maxReviewsPerCompany": 40
}
```

#### Example — weekly monitor (NEW / UPDATED only)

```json
{
  "startUrls": [{ "url": "https://clutch.co/agencies/digital-marketing" }],
  "incrementalMode": true,
  "monitorBaselineOnly": false,
  "maxPages": 10,
  "excludeSponsored": true
}
```

### Output

Each default row is a **company**. Reviews are nested on `reviews` when enabled (optional flat `type=review` rows via `emitReviewRows`).

| Field | Description |
| :--- | :--- |
| `providerId` | Stable Clutch pid (`data-clutch-pid`). |
| `name` / `profileUrl` / `websiteUrl` | Identity and decoded website. |
| `rating` / `reviewCount` | Aggregate score and volume. |
| `hourlyRate` / `minProjectSize` / `employees` | Published commercial bands. |
| `location` / `phone` / `address` | Geography and listing phone (not enriched email). |
| `services` | `{ name, percent }` mix from the card/profile. |
| `isVerified` / `isSponsored` / `listingPosition` | Trust + paid vs organic. |
| `reviews[]` | Title, rating, metrics, project size/length, reviewer industry/location/size. |

```json
{
  "type": "company",
  "providerId": "23730",
  "name": "Lounge Lizard",
  "profileUrl": "https://clutch.co/profile/lounge-lizard",
  "websiteUrl": "https://www.loungelizard.com/",
  "rating": 4.8,
  "reviewCount": 43,
  "minProjectSize": "$25,000+",
  "hourlyRate": null,
  "employees": "50 - 249",
  "location": "New York, NY",
  "isVerified": true,
  "isSponsored": true,
  "listingPosition": 1,
  "services": [
    { "name": "Web Development", "percent": 45 },
    { "name": "Web Design", "percent": 45 }
  ]
}
```

### Integration examples

#### Node.js

```javascript
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: process.env.APIFY_TOKEN });
const run = await client.actor('crawloop/clutch-scraper').call({
  startUrls: [{ url: 'https://clutch.co/web-developers' }],
  maxItems: 25,
  excludeSponsored: true,
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
console.log(items);
```

#### Python

```python
from apify_client import ApifyClient

client = ApifyClient(token="YOUR_TOKEN")
run = client.actor("crawloop/clutch-scraper").call(
    run_input={
        "startUrls": [{"url": "https://clutch.co/web-developers"}],
        "maxItems": 25,
        "excludeSponsored": True,
    }
)
items = client.dataset(run["defaultDatasetId"]).list_items().items
print(items)
```

#### cURL

```bash
curl "https://api.apify.com/v2/acts/crawloop~clutch-scraper/runs?token=YOUR_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"startUrls":[{"url":"https://clutch.co/web-developers"}],"maxItems":25}'
```

### MCP and AI assistants

Use this Actor from AI tools via [Apify MCP](https://docs.apify.com/platform/integrations/mcp).
Connect your Apify account, then call this Actor by its Store ID `crawloop/clutch-scraper`.

Example prompts:

- "Run Clutch Scraper for https://clutch.co/web-developers/new-york, exclude sponsored listings, min rating 4.5, and return the top 20 agencies as JSON"
- "Scrape the Clutch profile for lounge-lizard with reviews and summarize project sizes and quality scores — do not look up emails"
- "Chain Clutch Scraper then [Trustpilot Scraper](https://apify.com/crawloop/trustpilot-scraper) on the decoded websites for public reputation"
- "Run Clutch Scraper in monitor mode on https://clutch.co/agencies/digital-marketing and return only NEW or UPDATED agencies"

### Suite next step

After a Clutch shortlist, run [Trustpilot Scraper](https://apify.com/crawloop/trustpilot-scraper) on the same company websites for public reviews, or [Glassdoor Scraper](https://apify.com/crawloop/glassdoor-scraper) for employer/team signals before an RFP.

### FAQ

**Is this a Clutch API?**\
No. Clutch does not offer a public directory API. This Actor is a **Clutch API alternative** that reads public listing and profile HTML.

**Why not emails?**\
Agency BD and procurement already get websites, phones published on Clutch, and buyer context from reviews. Third-party inbox hunting is noisy, extra-billed, and not this product.

**How many companies per page?**\
A directory page currently renders on the order of **50–80 cards** (including sponsors). Pagination is `?page=N`. Out-of-range pages are clamped — the Actor stops when provider ids stop changing.

**Does it work with Python / Node.js / MCP?**\
Yes. Use the Apify client examples above or Apify MCP prompts.

### Related Actors

- [Trustpilot Scraper](https://apify.com/crawloop/trustpilot-scraper)
- [Glassdoor Scraper](https://apify.com/crawloop/glassdoor-scraper)
- [ThomasNet Scraper](https://apify.com/crawloop/thomasnet-scraper)
- [Kompass Scraper](https://apify.com/crawloop/kompass-scraper)

# Actor input Schema

## `startUrls` (type: `array`):

Clutch directory, location, search, or profile URLs. Example: https://clutch.co/web-developers or https://clutch.co/profile/lounge-lizard

## `category` (type: `string`):

Directory slug if you do not paste a URL (web-developers, agencies/digital-marketing, app-developers).

## `location` (type: `string`):

Country code (us, uk, in) or city/state slug (new-york, california). Combined with category as /us/web-developers or /web-developers/new-york.

## `fetchDetails` (type: `boolean`):

Open each company profile for address, year founded, extra locations, social links and service mix.

## `includeReviews` (type: `boolean`):

Attach client reviews (project size, ratings, reviewer industry/location). Nested on the company row. Does not find emails.

## `maxReviewsPerCompany` (type: `integer`):

Cap reviews collected per profile when includeReviews is on. 0 = all pages (still bounded by maxPages-style pagination).

## `emitReviewRows` (type: `boolean`):

Also push each review as its own dataset row (type=review) for table exports. Reviews stay nested on the company when includeReviews is on.

## `minRating` (type: `number`):

Keep companies with rating >= this value (procurement shortlist). Empty = no floor.

## `minReviews` (type: `integer`):

Keep companies with at least this many Clutch reviews.

## `verifiedOnly` (type: `boolean`):

Keep Clutch-verified companies only.

## `excludeSponsored` (type: `boolean`):

Drop paid/sponsor cards and keep organic directory ranking. Useful for competitor mapping.

## `maxItems` (type: `integer`):

Stop after this many company rows. 0 = unlimited (still bounded by maxPages).

## `maxPages` (type: `integer`):

Maximum listing pages per start URL. Clutch paginates with ?page=N and clamps overflowing pages — the Actor de-dupes by provider id.

## `incrementalMode` (type: `boolean`):

Compare this run to a named Key-Value Store. Emit NEW/UPDATED companies (rating, review count, rates, verified, rank). Skip UNCHANGED unless emitUnchanged is on. memo23 has no monitor.

## `monitorBaselineOnly` (type: `boolean`):

First scheduled run: seed KV fingerprints without emitting or billing company rows.

## `emitUnchanged` (type: `boolean`):

In monitor mode, also push UNCHANGED companies (still billed).

## `resetMonitorState` (type: `boolean`):

Clear MONITOR\_STATE in the named store before this run.

## `monitorStoreName` (type: `string`):

Named Key-Value Store for agency fingerprints (rating, reviews, hourly, rank).

## `compact` (type: `boolean`):

Drop long description and nested reviews (token-efficient for MCP).

## `proxyConfiguration` (type: `object`):

Apify Proxy. US residential is recommended — Clutch sits behind Cloudflare.

## Actor input object example

```json
{
  "startUrls": [
    {
      "url": "https://clutch.co/web-developers"
    }
  ],
  "fetchDetails": false,
  "includeReviews": false,
  "maxReviewsPerCompany": 50,
  "emitReviewRows": false,
  "verifiedOnly": false,
  "excludeSponsored": false,
  "maxItems": 50,
  "maxPages": 5,
  "incrementalMode": false,
  "monitorBaselineOnly": false,
  "emitUnchanged": false,
  "resetMonitorState": false,
  "monitorStoreName": "clutch-agency-monitor",
  "compact": false,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "US"
  }
}
```

# Actor output Schema

## `results` (type: `string`):

Default dataset items.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        {
            "url": "https://clutch.co/web-developers"
        }
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("crawloop/clutch-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "startUrls": [{ "url": "https://clutch.co/web-developers" }] }

# Run the Actor and wait for it to finish
run = client.actor("crawloop/clutch-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    {
      "url": "https://clutch.co/web-developers"
    }
  ]
}' |
apify call crawloop/clutch-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,crawloop/clutch-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/4xWX4NV33nHkOrVAt/builds/6eb84HPA1oaDSEZJx/openapi.json
