# Career Page Jobs Scraper: Greenhouse, Lever, Ashby, Workday (`datafetch_labs/ats-jobs-scraper`) Actor

Get every open job straight from company career pages. Type company names or URLs; it finds their board on Greenhouse, Lever, Ashby, Workday, Workable, SmartRecruiters, Recruitee, Personio or Breezy. Salary, location, remote flag, full description. Monitor mode: new jobs only.

- **URL**: https://apify.com/datafetch\_labs/ats-jobs-scraper.md
- **Developed by:** [DataFetch Labs](https://apify.com/datafetch_labs) (community)
- **Categories:** Open source
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.00 / 1,000 job scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Company Career Page Jobs Scraper

**Get every open job straight from company career pages: Greenhouse, Lever, Ashby, Workday, Workable, SmartRecruiters, Recruitee, Personio and Breezy HR, in one clean format.** Type in company names ("Stripe", "Airbnb") or URLs. The Actor finds each company's job board on its own and returns normalized jobs with salary, location, remote flag and full description. Turn on **monitor mode** to get only *new* jobs on every scheduled run.

- ✅ **Just type company names.** No need to know which ATS a company uses.
- ✅ **9 applicant tracking systems** under one output schema.
- ✅ **Straight from the source.** Jobs come from the employer's own board, not a stale aggregator, so they're fresh, complete and include the real apply link.
- ✅ **Structured salary** (`salaryMin`, `salaryMax`, currency, period) wherever the employer publishes it (Ashby, Lever, Recruitee, and pay-transparency postings).
- ✅ **Monitor mode** returns only jobs that appeared since the last run. Good for alerts, lead lists and hiring-signal tracking.
- ✅ **Fast and reliable.** Uses the official public job-board APIs: no browser, no proxies, no logins, no captchas.
- 💲 **$1 per 1,000 jobs.** You pay only for jobs returned.

### What can I use it for?

- **Job boards and job aggregators**: fill your board with fresh listings from hundreds of target companies.
- **Sales and recruiting agencies**: companies that are hiring have budget. Watch hiring surges by department (e.g. a company opening 10 sales roles).
- **Job seekers and career coaches**: follow dream companies and get new roles the day they're posted.
- **Market and salary research**: compare pay ranges, remote policies and team growth across companies.
- **AI agents and RAG**: feed LLMs up-to-date, structured job data with full descriptions.

### Supported job boards

| ATS | Example companies | How to pass it |
|---|---|---|
| Greenhouse | Stripe, Airbnb, Anthropic, Figma, Coinbase, Cloudflare | `Stripe` or `https://boards.greenhouse.io/stripe` or `greenhouse:stripe` |
| Lever | Spotify, Palantir | `Palantir` or `https://jobs.lever.co/palantir` |
| Ashby | OpenAI, Notion, Ramp, Linear, Zapier, Plaid | `OpenAI` or `https://jobs.ashbyhq.com/openai` |
| Workday | NVIDIA, Salesforce, and most large enterprises | Career site URL, e.g. `https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite` |
| Workable | Hugging Face | `huggingface` or `https://apply.workable.com/huggingface` |
| SmartRecruiters | ServiceNow, Canva, Bosch, Visa | `ServiceNow` or `https://jobs.smartrecruiters.com/ServiceNow` |
| Recruitee | Channable, bunq | `channable` or `https://channable.recruitee.com` |
| Personio | Many European SMBs | `personio` or `https://personio.jobs.personio.de` |
| Breezy HR | Many startups and SMBs | `https://company.breezy.hr` |

**How company lookup works:** a URL that points at a job board is used directly. Any other website (e.g. `notion.so` or `https://acme.com/careers`) is scanned for links to a supported ATS. A plain name is matched against the public boards of every supported ATS. The key-value store record `COMPANIES_REPORT` shows how each company was resolved, how many jobs its board has, and any errors.

### Input example

```json
{
    "companies": ["Stripe", "OpenAI", "notion.so", "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite"],
    "keywords": ["engineer", "developer"],
    "excludeKeywords": ["intern"],
    "locations": ["United States", "remote"],
    "postedWithinDays": 14,
    "includeDescription": true
}
```

All filters are optional. Leave them empty to get every open job.

### Output example

```json
{
    "id": "240d459b-696d-43eb-8497-fab3e56ecd9b",
    "ats": "ashby",
    "companySlug": "openai",
    "companyName": "OpenAI",
    "title": "Research Engineer",
    "department": "Research",
    "team": "Research",
    "location": "San Francisco",
    "locations": ["San Francisco"],
    "country": "United States",
    "isRemote": false,
    "workplaceType": null,
    "employmentType": "FullTime",
    "salaryMin": 250000,
    "salaryMax": 445000,
    "salaryCurrency": "USD",
    "salaryPeriod": "1 YEAR",
    "salaryText": "$250K – $445K • Offers Equity",
    "postedAt": "2025-04-05T00:03:20.653Z",
    "updatedAt": null,
    "url": "https://jobs.ashbyhq.com/openai/240d459b-696d-43eb-8497-fab3e56ecd9b",
    "applyUrl": "https://jobs.ashbyhq.com/openai/240d459b-696d-43eb-8497-fab3e56ecd9b/application",
    "descriptionText": "By applying to this role, you will be considered for Research Engineer roles across all teams at OpenAI...",
    "descriptionHtml": "<p>By applying to this role...</p>",
    "scrapedAt": "2026-09-28T20:34:06.492Z"
}
```

Every job has the same fields whatever its source ATS. A field is `null` when the employer doesn't publish it.

### Monitor mode: get only new jobs

1. Turn on **Monitor mode: only new jobs** and give the monitor a name (e.g. `eng-jobs-berlin`).
2. Run once. All matching jobs are returned and saved as the baseline.
3. [Schedule](https://docs.apify.com/platform/schedules) the Actor daily or hourly. Each run returns only jobs that weren't there before.
4. Hook it up to Slack, email, Google Sheets, Zapier, Make or n8n through Apify [integrations](https://docs.apify.com/platform/integrations) or webhooks.

Because runs that find nothing new return no items, a daily monitor of 100 companies usually costs only a few cents.

### Pricing

This Actor uses **pay per event** pricing: **$0.001 per job returned ($1 per 1,000 jobs)**. Filters run before billing, so jobs you filter out are free. Set a *maximum cost per run* and the Actor stops cleanly when it's reached.

Examples:

- Every open job at 50 mid-size tech companies (~5,000 jobs): about **$5**
- Daily monitor of 200 companies for new engineering roles (~30 new jobs/day): about **$0.90 per month**

### Tips

- **Faster runs**: turn off *Include full job description* if you only need titles, locations and links.
- **Workday**: pass the career site URL (open the company's job search and copy the URL, e.g. `https://company.wd1.myworkdayjobs.com/External`).
- **Company not found?** Pass the careers page URL instead of the name. If the careers page links to a supported ATS, it will be detected.
- **Big lists**: thousands of companies per run are fine. Raise *Companies processed in parallel* to go faster.

### FAQ

**Is this legal?** The Actor only reads public job postings that employers publish so they can be found, through the ATS vendors' public job board APIs. It collects no personal data. You're still responsible for how you use the data.

**Why is a company missing?** Some large companies (e.g. Amazon, Google, Meta) run fully custom career sites that aren't on a supported ATS. Open an issue with the company name and we'll look at adding it.

**Can I request another ATS?** Yes. Open an issue in the Issues tab. iCIMS, Taleo, Teamtailor, JazzHR and BambooHR are on the roadmap.

# Actor input Schema

## `companies` (type: `array`):

One company per line. Accepts a company name ("Stripe"), a website ("notion.so"), a careers page URL, an ATS board URL ("https://jobs.lever.co/spotify", "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite") or explicit "ats:slug" ("greenhouse:airbnb", "ashby:openai"). Workday boards must be given as a URL.

## `keywords` (type: `array`):

Keep only jobs whose title contains at least one of these (case-insensitive). Leave empty for all jobs.

## `excludeKeywords` (type: `array`):

Drop jobs whose title contains any of these, e.g. "intern", "senior".

## `locations` (type: `array`):

Keep only jobs whose location contains one of these (city, state or country, e.g. "London", "United States", "Germany"). Add "remote" to also include remote jobs.

## `remoteOnly` (type: `boolean`):

Keep only jobs marked remote by the employer or mentioning remote in the location.

## `departments` (type: `array`):

Keep only jobs whose department or team contains one of these, e.g. "Engineering", "Sales".

## `postedWithinDays` (type: `integer`):

Keep only jobs posted in the last N days. 0 = no limit. Jobs without a posting date are kept.

## `includeDescription` (type: `boolean`):

Adds the description as plain text and HTML. Turn off for faster runs when you only need titles and links.

## `onlyNewJobs` (type: `boolean`):

Remember jobs already seen and return only jobs that appeared since the previous run. Schedule the Actor (e.g. daily) to get a feed of new openings. The first run returns all matching jobs and saves them as the baseline.

## `monitorStateName` (type: `string`):

Name of the key-value store where seen jobs are remembered. Use a different name for each separate monitor (e.g. "eng-jobs-berlin").

## `maxJobsPerCompany` (type: `integer`):

0 = no limit.

## `maxJobs` (type: `integer`):

Stop after this many jobs. 0 = no limit.

## `maxConcurrency` (type: `integer`):

Companies processed in parallel.

## Actor input object example

```json
{
  "companies": [
    "Stripe",
    "https://jobs.ashbyhq.com/openai",
    "notion.so"
  ],
  "keywords": [],
  "remoteOnly": false,
  "postedWithinDays": 0,
  "includeDescription": true,
  "onlyNewJobs": false,
  "monitorStateName": "ats-jobs-monitor",
  "maxJobsPerCompany": 20,
  "maxJobs": 0,
  "maxConcurrency": 5
}
```

# Actor output Schema

## `results` (type: `string`):

No description

## `companiesReport` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "companies": [
        "Stripe",
        "https://jobs.ashbyhq.com/openai",
        "notion.so"
    ],
    "keywords": [],
    "maxJobsPerCompany": 20
};

// Run the Actor and wait for it to finish
const run = await client.actor("datafetch_labs/ats-jobs-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "companies": [
        "Stripe",
        "https://jobs.ashbyhq.com/openai",
        "notion.so",
    ],
    "keywords": [],
    "maxJobsPerCompany": 20,
}

# Run the Actor and wait for it to finish
run = client.actor("datafetch_labs/ats-jobs-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "companies": [
    "Stripe",
    "https://jobs.ashbyhq.com/openai",
    "notion.so"
  ],
  "keywords": [],
  "maxJobsPerCompany": 20
}' |
apify call datafetch_labs/ats-jobs-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,datafetch_labs/ats-jobs-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/FlcgxEx0vebiDV2Do/builds/XTqDZcCGwKzQurFJV/openapi.json
