# Career Site Jobs Scraper - Greenhouse, Lever, Ashby, Workable (`unbrowseai/career-site-jobs-scraper`) Actor

All open jobs from company career sites on Greenhouse, Lever, Ashby, Workable, SmartRecruiters and Recruitee. Paste a board, careers page or company name. Title, department, locations, remote flag, dates, apply link, description and parsed salary. Half the usual price.

- **URL**: https://apify.com/unbrowseai/career-site-jobs-scraper.md
- **Developed by:** [Unbrowse AI](https://apify.com/unbrowseai) (community)
- **Categories:** Jobs, Lead generation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.25 / 1,000 jobs

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Career Site Jobs Scraper – Greenhouse, Lever, Ashby, Workable, SmartRecruiters, Recruitee

Pull every open job straight from a company's own career site and get it back as clean, uniform data. Paste a job board link, a company website or careers page, or just a board name like `airbnb`. The scraper works out which applicant tracking system (ATS) the company uses, reads the board's public job feed and returns one row per job with the same fields no matter which system the job came from.

Supported job boards: **Greenhouse** (including EU boards and embedded boards), **Lever** (including EU), **Ashby**, **Workable**, **SmartRecruiters** and **Recruitee**.

### Why use it

- **Fresh data from the source.** Jobs are read live from the company's board at run time, not from a stale copy.
- **One schema for six systems.** Title, department, team, locations, remote/hybrid/onsite, employment type, posted and updated dates, job link, apply link and description all land in the same fields.
- **Salary included.** Pay ranges come from the board's own compensation fields (Greenhouse pay transparency, Ashby compensation, Lever salary range, Recruitee salary) and, when a board has none, from the pay range written in the description. Each salary says where it came from (`source: "ats"` or `"description"`) and its period (year, month, hour).
- **Works from a homepage.** Give it `https://linear.app/careers` or `https://www.figma.com/careers/` and it finds the board behind the page.
- **Pay only for matching jobs.** Filters run before anything is saved, and companies that fail or have no matches cost nothing.

### Use cases

- Sales prospecting: find companies hiring for roles your product serves.
- Recruiting and talent intelligence: track competitors' open roles, locations and pay bands.
- Job boards and aggregators: feed a niche board from a list of target companies.
- Market research: hiring velocity, remote share and salary ranges by department.

### How to use it

1. Add **career site URLs** (board links or company careers pages) and/or **company board names**. Pin a name to one system with a prefix such as `greenhouse:stripe` or `smartrecruiters:BoschGroup` (SmartRecruiters IDs are case-sensitive).
2. Optionally filter by **title keywords**, **excluded keywords**, **location**, **department**, **remote only** or **posted within N days**.
3. Set **max jobs per company** and, if you like, a **max for the whole run**.
4. Run it and download the results as JSON, CSV or Excel, or read them through the API. Schedule it daily to watch for new roles.

### Output example

```json
{
  "company": "Ramp",
  "ats": "ashby",
  "boardSlug": "ramp",
  "jobId": "34413f8d-26bf-4bbc-8ade-eb309a0e2245",
  "title": "Security Engineer, Cloud",
  "department": "Engineering",
  "team": "Backend",
  "location": "New York, NY (HQ)",
  "locations": ["New York, NY (HQ)", "Remote (Canada)", "Remote (US)", "Miami, FL"],
  "workplaceType": "hybrid",
  "isRemote": false,
  "employmentType": "Full-time",
  "postedAt": "2026-04-07T17:12:35.753Z",
  "updatedAt": null,
  "url": "https://jobs.ashbyhq.com/ramp/34413f8d-26bf-4bbc-8ade-eb309a0e2245",
  "applyUrl": "https://jobs.ashbyhq.com/ramp/34413f8d-26bf-4bbc-8ade-eb309a0e2245/application",
  "salary": { "min": 211400, "max": 290600, "currency": "USD", "period": "year", "text": "$211.4K - $290.6K", "source": "ats" },
  "descriptionText": "ABOUT RAMP\n\nRamp is building the smart infrastructure for finance teams…",
  "boardUrl": "https://jobs.ashbyhq.com/ramp",
  "input": "https://jobs.ashbyhq.com/ramp"
}
```

Turn on **Include description HTML** to also get `descriptionHtml`. Companies that could not be read appear as rows with `error` and `errorMessage` and are not charged.

### Pricing

One price per job saved:

| Apify plan | Price per 1,000 jobs |
|---|---|
| Free | $2.00 |
| Starter | $2.00 |
| Scale | $1.75 |
| Business and Enterprise | $1.25 |

About half of what comparable multi-board job scrapers charge per job. A small fee applies to each run start. Set a maximum cost per run and the scraper stops cleanly at that limit.

### FAQ

**My company isn't found by name.** Board names are not always the company name (for example `Linear` on Ashby). Paste the careers page URL instead and the board is detected from the page.

**Why is salary empty for some jobs?** Many employers publish no pay range. When neither the board nor the description states one, `salary` is `null` rather than a guess. Benefit amounts such as learning budgets are ignored.

**Does it support Workday, iCIMS or Taleo?** Not yet. Open an issue with the boards you need.

**Is it legal?** It reads job postings that companies publish openly for applicants. You are responsible for how you use the data, including the job board's terms and data protection law.

Missing a field or a job board? Tell us on the Issues tab.

# Actor input Schema

## `startUrls` (type: `array`):

Job board URLs (boards.greenhouse.io/airbnb, job-boards.greenhouse.io/vercel, jobs.lever.co/palantir, jobs.ashbyhq.com/ramp, apply.workable.com/huggingface, careers.smartrecruiters.com/BoschGroup, bunq.recruitee.com) or a company's own website or careers page (e.g. https://linear.app/careers): the job board it uses is found automatically.

## `companies` (type: `array`):

Optional. Board names as they appear in the job board address, e.g. stripe, notion, figma. Each is looked up on every supported job board. Pin one with a prefix: greenhouse:stripe, lever:palantir, ashby:ramp, workable:huggingface, smartrecruiters:BoschGroup, recruitee:bunq.

## `keywords` (type: `array`):

Keep jobs whose title contains any of these words (case-insensitive), e.g. engineer, designer. Filtered jobs are never charged.

## `excludeKeywords` (type: `array`):

Drop jobs whose title contains any of these words, e.g. intern, senior.

## `location` (type: `string`):

Keep jobs where any location contains this text, e.g. London, United States, Remote. Separate several with commas.

## `department` (type: `string`):

Keep jobs whose department or team contains this text, e.g. Engineering, Sales.

## `remoteOnly` (type: `boolean`):

Keep only jobs marked fully remote.

## `postedWithinDays` (type: `integer`):

Keep jobs first published in the last N days. Leave empty for all open jobs.

## `maxJobsPerCompany` (type: `integer`):

Stop after this many matching jobs from one company.

## `maxJobs` (type: `integer`):

Optional ceiling for the whole run, across all companies.

## `includeDescription` (type: `boolean`):

Add the full description as plain text. Salary is also read from the description when the job board has no pay field.

## `includeHtml` (type: `boolean`):

Also add the description as the original HTML (descriptionHtml).

## Actor input object example

```json
{
  "startUrls": [
    "https://boards.greenhouse.io/airbnb",
    "https://jobs.ashbyhq.com/ramp"
  ],
  "companies": [],
  "keywords": [],
  "excludeKeywords": [],
  "remoteOnly": false,
  "maxJobsPerCompany": 10,
  "includeDescription": true,
  "includeHtml": false
}
```

# Actor output Schema

## `jobs` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        "https://boards.greenhouse.io/airbnb",
        "https://jobs.ashbyhq.com/ramp"
    ],
    "maxJobsPerCompany": 10
};

// Run the Actor and wait for it to finish
const run = await client.actor("unbrowseai/career-site-jobs-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "startUrls": [
        "https://boards.greenhouse.io/airbnb",
        "https://jobs.ashbyhq.com/ramp",
    ],
    "maxJobsPerCompany": 10,
}

# Run the Actor and wait for it to finish
run = client.actor("unbrowseai/career-site-jobs-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    "https://boards.greenhouse.io/airbnb",
    "https://jobs.ashbyhq.com/ramp"
  ],
  "maxJobsPerCompany": 10
}' |
apify call unbrowseai/career-site-jobs-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,unbrowseai/career-site-jobs-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/QLsa6NIfeh0Nbg4uA/builds/IrRnrLqFIzufzIfq4/openapi.json
