# Company Jobs Scraper – Greenhouse, Lever, Ashby & Workday (`anatolia-data/ats-jobs-scraper`) Actor

Every open job from any company's Greenhouse, Lever, Ashby or Workday career site in one clean format: title, team, location, remote, salary, posting date and full description. Paste a board link, a company name or just the company website.

- **URL**: https://apify.com/anatolia-data/ats-jobs-scraper.md
- **Developed by:** [Oğulcan Katmer](https://apify.com/anatolia-data) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.50 / 1,000 jobs

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Company Jobs Scraper – Greenhouse, Lever, Ashby & Workday

Get every open job from any company's career site in one clean format — whether the company uses Greenhouse, Lever, Ashby or Workday.

Paste a job board link, a board name, or **just the company website**: the scraper finds the company's job board for you. No login, no API key. It reads the same public job feeds the companies' own career pages use.

### What you get

#### Per job — the same fields on every platform

| Field | Description |
|---|---|
| `title`, `department`, `team` | What the role is and where it sits |
| `location`, `locations`, `country` | Every location the job is open in |
| `workplaceType` | `remote`, `hybrid` or `onsite` |
| `employmentType` | Full-time, part-time, contract… |
| `salary` | `{ min, max, currency, interval, source }` — from the platform's own pay field, or read out of the description where pay-transparency laws put it |
| `postedAt` | When the job went live (Workday only gives relative dates like "Posted 30+ Days Ago"; `postedAtIsApproximate` flags those) |
| `url`, `applyUrl` | The job page and the application form |
| `description` | Full description as clean text, with bullet points kept |
| `company`, `platform`, `jobId`, `requisitionId` | Identifiers |

#### Per company — a free hiring summary

- Open jobs right now
- New jobs in the last 7 and 30 days — a direct **hiring-intent signal**
- Share of remote or hybrid roles
- Top departments, locations and countries
- How many jobs publish a salary

### Input

```json
{
  "companies": [
    "https://boards.greenhouse.io/airbnb",
    "https://jobs.lever.co/palantir",
    "https://jobs.ashbyhq.com/ramp",
    "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite",
    "linear.app",
    "figma"
  ],
  "keywords": ["engineer", "data"],
  "postedWithinDays": 30
}
```

Three ways to name a company:

- **Board link** — the fastest and most precise. Required for Workday.
- **Board name** — `airbnb`. Greenhouse, Lever and Ashby are all checked.
- **Company website** — `linear.app`. The homepage and careers page are read to find the board link.

#### Filters

`keywords` (title, department and team — add `searchDescriptions` to match the full text), `locations`, `departments`, `remoteOnly`, `postedWithinDays`.

### Common uses

- **Job boards and aggregators** — one feed across four platforms, deduplicated and normalized
- **Sales prospecting** — companies hiring for a role are companies about to buy tools for that team
- **Recruiting and sourcing** — watch target companies for new openings every day
- **Market research** — hiring velocity, remote share and pay ranges across a list of companies

### Notes and limits

- Only publicly listed jobs are returned. Internal postings are not in the public feed.
- Workday lists at most 2,000 jobs per search. `maxJobsPerCompany` caps the rest; the summary says when a board was read only partly.
- Salaries appear only where the company publishes them.
- Independent tool. Not affiliated with Greenhouse, Lever, Ashby or Workday.

### Pricing

Pay per job returned. Company summaries and status records are free. You can cap spend per run with **Max cost per run** in the Apify Console.

### Support

Something missing or wrong? Open an issue on the **Issues** tab with the input you used and the run ID.

# Actor input Schema

## `companies` (type: `array`):

One per line. Any of: a job board link (boards.greenhouse.io/airbnb, jobs.lever.co/palantir, jobs.ashbyhq.com/ramp, nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite), a board name (airbnb) or the company website (airbnb.com) — the board is found for you.

## `startUrls` (type: `array`):

Same as above, for long lists loaded from a file or another Actor.

## `keywords` (type: `array`):

Only jobs whose title, department or team contains one of these words (case-insensitive). Leave empty for every open job.

## `searchDescriptions` (type: `boolean`):

Match keywords against the full job description as well, not only the title.

## `locations` (type: `array`):

Only jobs in these places — a city, state or country as the company writes it (London, Germany, CA). Case-insensitive.

## `departments` (type: `array`):

Only jobs in these departments or teams, for example Engineering, Sales.

## `remoteOnly` (type: `boolean`):

Only jobs marked remote or hybrid.

## `postedWithinDays` (type: `integer`):

Only recently posted jobs — the fastest way to track who is hiring right now.

## `maxJobsPerCompany` (type: `integer`):

Stop reading a company's board after this many jobs.

## `includeDescription` (type: `boolean`):

Adds the full description as plain text. For Workday this needs one extra request per job.

## `includeCompanySummary` (type: `boolean`):

One free record per company: open jobs, new jobs in the last 7 and 30 days, remote share, top departments and locations.

## `proxyConfiguration` (type: `object`):

Apify Proxy is recommended for long lists.

## `maxConcurrency` (type: `integer`):

How many requests run at once.

## Actor input object example

```json
{
  "companies": [
    "https://boards.greenhouse.io/airbnb",
    "https://jobs.ashbyhq.com/ramp",
    "linear.app"
  ],
  "startUrls": [],
  "searchDescriptions": false,
  "remoteOnly": false,
  "maxJobsPerCompany": 1000,
  "includeDescription": true,
  "includeCompanySummary": true,
  "proxyConfiguration": {
    "useApifyProxy": true
  },
  "maxConcurrency": 10
}
```

# Actor output Schema

## `results` (type: `string`):

Every job, followed by the company summaries.

## `overview` (type: `string`):

Open the dataset in a table view.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "companies": [
        "https://boards.greenhouse.io/airbnb",
        "https://jobs.ashbyhq.com/ramp",
        "linear.app"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("anatolia-data/ats-jobs-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "companies": [
        "https://boards.greenhouse.io/airbnb",
        "https://jobs.ashbyhq.com/ramp",
        "linear.app",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("anatolia-data/ats-jobs-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "companies": [
    "https://boards.greenhouse.io/airbnb",
    "https://jobs.ashbyhq.com/ramp",
    "linear.app"
  ]
}' |
apify call anatolia-data/ats-jobs-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,anatolia-data/ats-jobs-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/auVyfUwkG7ZoJGf2W/builds/8NVWffYEeFWICb6OJ/openapi.json
