# Chamber of Commerce Directory Scraper: Local Business Lists (`ledgerstar/chamber-directory-scraper`) Actor

Extract chamber of commerce member directories into clean business lists. Supports GrowthZone, ChamberMaster and schema.org sites. Includes business name, address, city, ZIP, category, phone and website. Fetches pages politely. No personal data. Ready for accountants, agencies and chamber staff.

- **URL**: https://apify.com/ledgerstar/chamber-directory-scraper.md
- **Developed by:** [Ledger Star](https://apify.com/ledgerstar) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$5.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Chamber of Commerce Directory Scraper: Local Business Lists

Extract local businesses from any public chamber of commerce member directory. This Actor turns GrowthZone and ChamberMaster directories (and any directory tagged with schema.org business markup) into a clean spreadsheet: business name, address, city, state, ZIP, category and phone number.

### What you get

- Each record has business name, street address, city, region, ZIP, phone and website. Optional fields include description, full category list and social profiles (Facebook, Instagram, LinkedIn company page, X).
- Works with thousands of directories hosted on GrowthZone and ChamberMaster platforms (URLs containing /list/) and any directory tagged with schema.org LocalBusiness or Organization markup.
- Follows category pages automatically to capture all members, deduplicates across categories on the same directory, respects each site's rules for automated access and fetches pages politely, and never collects personal names or email addresses.

### Quick start

1. Find a public chamber of commerce directory URL containing /list/ (for example, https://business.yourchamber.org/list/).
2. Paste the directory URL and set a maximum results limit. Optionally turn on member pages to extract descriptions and social links.
3. Click Start, then export the results.

**Worked example:** Colorado Springs Chamber of Commerce directory (https://business.coloradospringschamberedc.com/list/), maximum 20 results. This costs $0.10 at $5.00 per 1,000 results.

### Who uses it

- **Accountants and bookkeepers:** export a chamber directory filtered by industry category to find established businesses that may need tax services, bookkeeping or payroll.

- **Agencies and freelancers selling web, marketing or IT services:** export a chamber directory and filter for businesses with no website to identify companies that need what you sell.

- **Chambers, economic developers and researchers:** export your own directory for membership analysis, compare membership by category across regions, or keep a business map up to date with monthly scheduled runs.

### Sample output

Example rows showing the format:

| Name | City | Region | Phone | Website |
|---|---|---|---|---|
| Capitol CPA Group, LLC | Springfield | IL | (217) 555-0101 | https://capitolcpa.example/ |
| Main Street Dental | Springfield | IL | (217) 555-0102 | https://mainstreeidental.example/ |
| Premier Landscaping Services | Springfield | IL | (217) 555-0103 | https://landscaping.example/ |

<details><summary>Full JSON example</summary>

```json
{
  "name": "Capitol CPA Group, LLC",
  "street": "500 S 2nd St Suite 300",
  "city": "Springfield",
  "region": "IL",
  "postalCode": "62701",
  "phone": "(217) 555-0101",
  "website": "https://capitolcpa.example/",
  "directoryCategory": "Accounting",
  "categories": ["Accounting", "Tax Preparation"],
  "description": "Full-service CPA firm for small businesses: monthly bookkeeping, payroll, tax planning and returns.",
  "facebook": "https://www.facebook.com/capitolcpagroup",
  "instagram": null,
  "linkedin": "https://www.linkedin.com/company/capitol-cpa-group",
  "x": null,
  "memberUrl": "https://business.example.org/list/member/capitol-cpa-group-1002",
  "directoryUrl": "https://business.example.org/list/ql/accounting-3",
  "source": "Chamber of commerce member directory (public listing)",
  "scrapedAt": "2026-09-25T14:00:00.000Z"
}
```

</details>

### Input

| Field | Type | Default | Description |
|---|---|---|---|
| Directory URLs | array | none | Directory home (e.g. https://business.yourchamber.org/list/) or a single category page. Only public pages; the Actor never logs in. |
| Maximum results | integer | 20 | Stop after this many businesses. You are charged per result. |
| Follow category pages | boolean | true | From a directory home page, visit each category page on the same site. |
| Open each member page for description, categories and social links | boolean | false | One extra request per business (slower). Off by default. |
| Delay between requests (seconds) | number | 1.5 | Politeness delay per site. Minimum 1 second; site crawl-speed preferences are honoured if larger. |

<details><summary>Advanced options</summary>

- **Maximum listing pages** (integer, default 50): Safety limit on directory pages visited per run.
- **Only new since** (string, optional): Enter a date in YYYY-MM-DD format or the word lastRun to receive only new listings. Turns on alert mode.
- **Alert mode: lookback safety window in days** (integer, default 7, max 30): Only used with lastRun above. Checks this many days before your last run for listings the directory published late. You are never charged twice for the same listing.
- **Alert state name** (string, optional): Only used with alert mode. Give this a fixed name such as "my-chamber-nightly" so this schedule always uses the same saved state, even if you change other filters. Letters, numbers and hyphens only, up to 30 characters.

</details>

### Alert mode: only new records

Run this Actor on a schedule and get only listings new since your last run. The first run returns normal results; every run after that returns only new businesses. You pay only for new results, so a run with nothing new costs nothing.

Set it up in three steps:

1. Create a Task in Apify Console with all your filters (directory URLs, follow categories, member pages, delay).
2. In the Task input, set "Only new since" to lastRun.
3. Go to the Task Schedule tab, click Add Schedule and choose daily or weekly.

Keep your filters the same on every run. How the Actor remembers what it already sent is explained in the FAQ.

### Pricing

$5.00 per 1,000 results ($0.005 per business).

Worked examples: 100 results cost $0.50, 1,000 results cost $5.00, and 10,000 results cost $50.00. Set a maximum cost per run and the Actor stops cleanly at your budget.

### Integrations

- **Google Sheets:** send every run's results to a sheet automatically.
- **Zapier and Make:** trigger a workflow for each new listing, for example add it to your CRM.
- **Slack:** post a message when a run finds new businesses.
- **Apify API:** download results as CSV, Excel or JSON from your own code.
- **Export:** CSV, Excel and JSON downloads straight from the dataset view.

### Data source and compliance

Chamber member directories are published so the public can find local businesses. This Actor only visits publicly accessible pages, never logs in, respects each site's crawl policies and stays on the directory's own website. It collects business information only and never personal names or email addresses.

Each chamber sets its own website terms. Only run the Actor on directories whose terms allow it, and follow anti-spam rules (CAN-SPAM, TCPA and local equivalents) when contacting businesses. This Actor is not affiliated with GrowthZone, ChamberMaster or any chamber of commerce.

### FAQ

**How do I know if a directory is supported?**
If its URL contains /list/ (e.g. /list/, /list/ql/accounting, /list/member/…) it works out of the box. Other directories work if they use schema.org LocalBusiness or Organization markup. To find one, search for your city name plus "chamber of commerce member directory" and look for /list/ in the URL. Test with a small maximum results limit.

**Why did I get fewer results than the directory shows?**
Common reasons: a site blocks some pages from crawling (they are skipped); a category page lists businesses only by letter and you hit the page limit; or maximum results was reached. The run summary shows what happened.

**Can I get email addresses or owner names?**
No. This Actor is designed not to collect personal contact data.

**How does alert mode remember which businesses to skip?**
The Actor records the first time it encountered each business as that business's first-seen date. With lastRun you get only new listings; with a date you get listings first seen on or after that date. The Actor looks back up to 30 days before the last run to catch listings published late, and never charges twice for the same business. It saves the run timestamp and delivered ids to a named store called alert-chamber-directory-scraper-{code}. To start over, delete that store in Storage or set a custom alert state name.

**How fast is it?**
About one page per 1.5 seconds per site. A category page usually holds dozens of businesses, so most directories finish in a few minutes. Turning on member pages adds about 1.5 seconds per business.

### More from Ledgerstar

- **Texas New Business Leads:** Sales tax permits from the Texas Comptroller: https://apify.com/ledgerstar/texas-sales-tax-permits

- **New Business Filings:** Colorado and New York business formation data: https://apify.com/ledgerstar/state-business-filings

- **Nonprofit Finder:** IRS Exempt Organizations by state: https://apify.com/ledgerstar/nonprofit-finder

- **Building Permits Leads:** new construction and contractor permits across several cities (arriving on the Store soon)

### Support

Support: open an issue on this Actor's Issues tab in Apify Console.

# Actor input Schema

## `startUrls` (type: `array`):

A directory home (e.g. https://business.yourchamber.org/list/) or a single category page (…/list/ql/accounting-3). Only public pages; the Actor never logs in.

## `maxItems` (type: `integer`):

Stop after this many businesses. You are charged per result.

## `newSince` (type: `string`):

This directory has no dates on its own, so the Actor records the first run in which it saw each business as that business's first seen date. Leave empty for normal results. Set an ISO date (YYYY-MM-DD) to get only businesses first seen on or after that date, or set the word lastRun to get only businesses that were not yet seen before the previous run (useful for a schedule).

## `alertLookbackDays` (type: `integer`):

Only used with the word lastRun above. It matters little for this Actor, since new here means first seen by this Actor rather than a source publish date, but it is still checked as a safety margin. You are never charged twice for the same business, so a larger number is always safe.

## `followCategoryPages` (type: `boolean`):

From a directory home page, visit each category page on the same site.

## `includeMemberDetails` (type: `boolean`):

One extra request per business (slower). Off by default.

## `maxPages` (type: `integer`):

Safety limit on directory pages visited per run.

## `requestDelaySecs` (type: `number`):

Politeness delay per site. Minimum 1 second; a robots.txt Crawl-delay is honoured if larger.

## `alertStateKey` (type: `string`):

Optional. Gives the newSince first seen tracking a stable name instead of one generated from your other input fields, so you can reuse it across schedules with different filters. Letters, numbers and hyphens only, up to 30 characters.

## Actor input object example

```json
{
  "startUrls": [
    {
      "url": "https://business.coloradospringschamberedc.com/list/"
    }
  ],
  "maxItems": 20,
  "alertLookbackDays": 7,
  "followCategoryPages": true,
  "includeMemberDetails": false,
  "maxPages": 50,
  "requestDelaySecs": 1.5
}
```

# Actor output Schema

## `results` (type: `string`):

No description

## `runSummary` (type: `string`):

No description

## `output` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        {
            "url": "https://business.coloradospringschamberedc.com/list/"
        }
    ],
    "maxItems": 20,
    "alertLookbackDays": 7
};

// Run the Actor and wait for it to finish
const run = await client.actor("ledgerstar/chamber-directory-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "startUrls": [{ "url": "https://business.coloradospringschamberedc.com/list/" }],
    "maxItems": 20,
    "alertLookbackDays": 7,
}

# Run the Actor and wait for it to finish
run = client.actor("ledgerstar/chamber-directory-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    {
      "url": "https://business.coloradospringschamberedc.com/list/"
    }
  ],
  "maxItems": 20,
  "alertLookbackDays": 7
}' |
apify call ledgerstar/chamber-directory-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,ledgerstar/chamber-directory-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/hL6z2vQgK0DRN9Uf3/builds/exD29o1rF14V0ZwK6/openapi.json
