# Career Site Jobs Scraper (`brightpath-data/career-site-jobs`) Actor

Live job postings from 10 applicant tracking systems by company name, domain or careers URL

- **URL**: https://apify.com/brightpath-data/career-site-jobs.md
- **Developed by:** [Nick Randall](https://apify.com/brightpath-data) (community)
- **Categories:** Jobs, Lead generation, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $3.00 / 1,000 job posting records

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Career Site Jobs Scraper

Live job postings from 10 applicant tracking systems, in one call, by company name, domain or careers URL.

Get clean, structured job listings from Greenhouse, Lever, Ashby, SmartRecruiters, Workable, Recruitee, Personio, Teamtailor, BambooHR and Breezy as JSON, CSV or Excel, or call it as a tool from Claude, Cursor, ChatGPT or any MCP client. Pay only for the results you receive.

### What you get

Give it a list of companies (bare names, domains, or a careers URL you copied from a browser tab) and it returns every open job it can find across the 10 supported ATS platforms: title, department, location, remote flag, employment type, posted and updated dates, the job's own URL, and a plain-text description (where the ATS's list endpoint includes one).

A recognized careers URL (for example `https://boards.greenhouse.io/stripe` or `https://jobs.lever.co/some-company`) is routed straight to its ATS. A bare company name or domain is tried against each ATS in turn until one has a live board, so you do not need to know which system a company uses.

Every result is a flat record with stable field names, so it drops straight into a spreadsheet, a database or an AI agent's context.

### Why use this instead of the website

- One Actor covers 10 ATS platforms instead of a separate tool per vendor
- Filters and pagination handled for you, with automatic retries and polite rate limiting
- Results in JSON, CSV, Excel or via API, or piped into Zapier, Make, n8n and Google Sheets
- Works as an MCP tool, so AI agents can fetch open roles on demand
- No browser, no proxies, no personal data: fast runs and a tiny cost per result

### Input

| Field | Type | Default | Meaning |
|-------|------|---------|---------|
| `companies` | array | | Company names, domains or careers URLs, one per entry |
| `atsList` | array | all 10 | ATS platforms to try for a bare company name or domain |
| `postedAfter` | string | | ISO date; only jobs first posted on or after this date |
| `updatedAfter` | string | | ISO date; only jobs posted or last updated on or after this date |
| `maxResults` | integer | 200 | Cap on results saved. You are charged per result, so this caps your cost. |

Example input:

```json
{
  "companies": ["stripe", "https://boards.greenhouse.io/airbnb"],
  "maxResults": 200
}
```

### Output

Example result:

```json
{
  "ats": "greenhouse",
  "company": "stripe",
  "jobId": "7010123456",
  "title": "Software Engineer, Payments",
  "department": "Engineering",
  "location": "Remote - US",
  "remote": true,
  "employmentType": null,
  "url": "https://job-boards.greenhouse.io/stripe/jobs/7010123456",
  "postedAt": "2026-08-15T00:00:00Z",
  "updatedAt": "2026-09-01T00:00:00Z",
  "description": "We are looking for a Software Engineer to join our Payments team..."
}
```

Field reference:

- `ats`: which of the 10 ATS platforms the job came from
- `company`: the company slug or name you supplied
- `jobId`, `title`, `department`, `location`, `remote`, `employmentType`
- `url`: the job's own listing page
- `postedAt`, `updatedAt`: ISO dates where the ATS provides them
- `description`: plain text (HTML tags stripped), null where the ATS's list endpoint does not include one (SmartRecruiters, Workable, BambooHR require a per-job call this Actor skips to keep cost low)

### Pricing

Pay per event. You are charged **$3.00 per 1,000 results** saved to the dataset, plus a fraction of a cent per run start. Nothing is charged for results you do not receive. Set "Max total charge per run" in the run options to cap spending on any run. When a run reaches your cap it stops cleanly and keeps everything it already saved.

Rough guide: 1,000 results cost $3.00 and take about 25 seconds.

### Use it from an AI agent (MCP)

This Actor is available as an MCP tool through the Apify MCP server. Add it to your client, then ask the agent for the data in plain language.

Claude Desktop, Claude Code or Cursor (`mcp.json` / `claude_desktop_config.json`):

```json
{
  "mcpServers": {
    "apify": {
      "url": "https://mcp.apify.com/?actors=brightpath-data/career-site-jobs",
      "headers": { "Authorization": "Bearer YOUR_APIFY_TOKEN" }
    }
  }
}
```

ChatGPT and other clients that support remote MCP servers: add `https://mcp.apify.com/?actors=brightpath-data/career-site-jobs` as a connector with your Apify token.

Example prompt once connected: "Find open engineering roles at Stripe and Airbnb posted in the last month."

### Use it from code

```bash
curl -X POST "https://api.apify.com/v2/acts/brightpath-data~career-site-jobs/run-sync-get-dataset-items?token=YOUR_APIFY_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"companies":["stripe","https://boards.greenhouse.io/airbnb"],"maxResults":200}'
```

Python:

```python
from apify_client import ApifyClient
client = ApifyClient("YOUR_APIFY_TOKEN")
run = client.actor("brightpath-data/career-site-jobs").call(run_input={"companies": ["stripe", "https://boards.greenhouse.io/airbnb"], "maxResults": 200})
items = client.dataset(run["defaultDatasetId"]).list_items().items
```

### Limits and fair use

- Up to 10,000 results per run.
- A bare company name is tried against every ATS in `atsList` (10 by default) until one matches; set `atsList` to the one you already know to cut requests and speed up large batches.
- SmartRecruiters is capped at 2,000 postings per company per run.
- Companies with no public board on any tried ATS return nothing for that entry rather than an error.
- All requests are paced and retried; a company behind an access challenge is skipped rather than bypassed.

### Data source and legal

Data comes from the public, unauthenticated job board endpoints that Greenhouse, Lever, Ashby, SmartRecruiters, Workable, Recruitee, Personio, Teamtailor, BambooHR and Breezy each publish for their customers' career sites; Lever's own documentation states postings may be read by third parties, and Personio additionally publishes a dedicated XML feed built for republishing. This Actor collects public, non-personal data only (no recruiter names or contact details) and does not bypass logins, paywalls or access controls. You are responsible for how you use the data.

### Support

Found a problem or need a field added? Open an issue on the Actor's Issues tab. Fixes for broken runs are prioritized.

# Actor input Schema

## `companies` (type: `array`):

Company names, domains or careers URLs, one per line (e.g. "stripe", "stripe.com", "https://boards.greenhouse.io/stripe"). A recognized careers URL is routed straight to its ATS; a bare name or domain is tried against every ATS in atsList until one matches.

## `atsList` (type: `array`):

Applicant tracking systems to try, in order, when a company entry is a bare name or domain rather than a recognized careers URL. Leave empty to try all 10.

## `postedAfter` (type: `string`):

ISO date (e.g. 2026-08-01). Only jobs first posted on or after this date are returned. Leave empty for no filter.

## `updatedAfter` (type: `string`):

ISO date. Only jobs posted or last updated on or after this date are returned. Leave empty for no filter.

## `maxResults` (type: `integer`):

Maximum number of jobs to save. You are charged per result saved, so this also caps the cost of a run.

## Actor input object example

```json
{
  "companies": [
    "stripe",
    "https://boards.greenhouse.io/airbnb"
  ],
  "atsList": [],
  "maxResults": 200
}
```

# Actor output Schema

## `results` (type: `string`):

The dataset with one flat record per result. Append ?format=csv or ?format=xlsx to the URL for other formats.

## `summary` (type: `string`):

OUTPUT record in the key-value store: counts of results pushed and charged, requests, retries and duration.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "companies": [
        "stripe",
        "https://boards.greenhouse.io/airbnb"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("brightpath-data/career-site-jobs").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "companies": [
        "stripe",
        "https://boards.greenhouse.io/airbnb",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("brightpath-data/career-site-jobs").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "companies": [
    "stripe",
    "https://boards.greenhouse.io/airbnb"
  ]
}' |
apify call brightpath-data/career-site-jobs --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,brightpath-data/career-site-jobs"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/MrXoAp8jzN46HAeqh/builds/JjfTm6wWYukxXoHjy/openapi.json
