Greenhouse Jobs Scraper avatar

Greenhouse Jobs Scraper

Pricing

Pay per event

Go to Apify Store
Greenhouse Jobs Scraper

Greenhouse Jobs Scraper

Greenhouse jobs scraper and API: export every open job listing from any company's Greenhouse job board (title, location, departments, offices, full description, publish dates) to JSON, CSV or Excel. Official public Job Board API, no login, keyword and location filters.

Pricing

Pay per event

Rating

0.0

(0)

Developer

COMPASSLAB

COMPASSLAB

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

4 hours ago

Last modified

Categories

Share

Get Greenhouse job listings as clean JSON, CSV or Excel, from the Apify API or on a schedule. No login, about $1.60 for 1,000 jobs, $4.00 with full descriptions.

What does Greenhouse Jobs Scraper do?

Greenhouse Jobs Scraper extracts structured data from job-boards.greenhouse.io. Collect all open jobs from public Greenhouse-hosted company job boards through the official Job Board API, with optional keyword and location filters, for job market monitoring, lead generation and recruiting research. It works as an API for job-boards.greenhouse.io data: run it from Apify Console, on a schedule, or from your own code, and get clean, typed JSON with numbers as numbers and dates in ISO 8601.

What you get

Data12 fields per item: jobId, title, companyName, location, departments, offices, ...
FormatsJSON, CSV, Excel, HTML, or the Apify API
Price$1.60 per 1,000 jobs, $4.00 with full descriptions, pay per result
AccessPublic job-boards.greenhouse.io data only: no login, no cookies, robots.txt respected

Why use Greenhouse Jobs Scraper?

  • Recruiters and staffing agencies: see which companies are hiring, for which roles and where, and reach out first.
  • Job aggregators and job boards: feed fresh, de-duplicated listings into your own board on a schedule.
  • Market research and HR analytics: track hiring trends, salaries (where published), locations and remote share.
  • AI agents and RAG: give an LLM live, structured job data instead of stale web pages.

Main features:

  • Gets each source's results in as few requests as possible and stops at maxItems results.
  • Filters: companies (companies), keywords (keywords), excludeKeywords (exclude keywords), location (location filter), remoteOnly (remote only), postedAfter (posted after), keyword (title keyword), includeDescription (include description), onlyNewSinceLastRun (only new jobs since the last run), so you only get (and pay for) the results you need.
  • Polite by default: respects robots.txt, at most maxConcurrency parallel requests and a delay between requests.
  • Checks every result against field validators, so layout changes show up as clear data-quality warnings.
  • Runs on the Apify platform: scheduling, API access, integrations, monitoring and datasets you can export.

What data can Greenhouse Jobs Scraper extract?

FieldTypeDescription
jobIdintegerGreenhouse job ID
titlestringJob title
companyNamestringBoard token identifying the company (e.g. airbnb)
locationstringJob location name as listed on the board
departmentsarrayNames of the departments the job belongs to
officesarrayNames of the offices the job belongs to
urlstringAbsolute URL of the job posting (absolute_url)
updatedAtISO 8601 dateLast update time (ISO 8601)
firstPublishedISO 8601 dateFirst publication time (ISO 8601)
requisitionIdstringInternal requisition ID if the company provides one
languagestringLanguage code of the posting (e.g. en)
descriptionTextstringJob description with HTML stripped, truncated to 5000 characters

How to scrape job-boards.greenhouse.io

  1. Open Greenhouse Jobs Scraper in Apify Console and go to the Input tab.
  2. Enter what to scrape (see the Input section below), for example the start URLs.
  3. Set Max items to the number of results you need.
  4. Click Start and wait for the run to finish.
  5. Download the results from the Output tab, or fetch them with the API.

How much will it cost to scrape job-boards.greenhouse.io?

This Actor is priced per result: $1.60 per 1,000 results, with no extra charge for platform usage. That is about $1.60 for 1,000 jobs, $4.00 with full descriptions: 100 results cost $0.16 and 10,000 results cost $16.00. Set a maximum cost per run and the Actor stops when it is reached. With Include description on, each job comes with its full description text and is priced $4.00 per 1,000 results (job with description). Filters (companies, keywords, location, remote only, posted after, only new jobs) run before charging: you never pay for jobs you filtered out.

Input

See the Input tab for full configuration options.

FieldTypeRequiredDescription
companiesarraynoCompany names (e.g. Stripe) or job board URLs. Each name is looked up on Greenhouse; names that can't be found are logged and skipped.
keywordsarraynoKeep jobs whose title (and description, when 'Include description' is on) contains any of these words. Case-insensitive.
excludeKeywordsarraynoDrop jobs whose title (and description, when on) contains any of these words.
locationstringnoOptional. Only return jobs whose location name contains this text (case-insensitive).
remoteOnlybooleannoKeep only remote jobs (location says Remote/Anywhere, or the source marks the job as remote).
postedAfterstringnoKeep jobs posted on or after this date (YYYY-MM-DD). Jobs without a date are kept.
keywordstringnoOptional. Only return jobs whose title contains this text (case-insensitive).
includeDescriptionbooleannoAdd the full job description text. Priced as 'job with description'.
onlyNewSinceLastRunbooleannoFor scheduled runs: skip jobs this same input already returned. The IDs are kept in a named key-value store; delete it to start over.
maxItemsintegernoMaximum number of items to return (0 = unlimited).
startUrlsarraynoGreenhouse board URLs (https://job-boards.greenhouse.io/{boardToken} or https://boards.greenhouse.io/{boardToken}) or boards-api.greenhouse.io API URLs. The board token is extracted from each URL.
maxPagesintegernoMaximum listing pages to follow per start URL (pagination).
maxConcurrencyintegernoMaximum parallel requests (politeness; 1-10).
requestDelayMsintegernoMinimum delay between requests, in milliseconds (at least 250).
proxyTypestringnonone (direct connection), datacenter (Apify Proxy, cheapest) or residential (opt-in, billed per GB, fewer blocks). The actor never switches by itself.
proxyCountrystringnoTwo-letter country code for the proxy IP (optional).

Example input:

{
"startUrls": [
{
"url": "https://job-boards.greenhouse.io/airbnb"
}
],
"maxItems": 75,
"maxPages": 3,
"maxConcurrency": 2,
"requestDelayMs": 1000,
"proxyType": "none",
"includeDescription": true
}

Output

You can download the dataset in various formats such as JSON, HTML, CSV, or Excel. Example results from a real run:

[
{
"jobId": 8232207,
"title": "Accountant",
"companyName": "airbnb",
"location": "United States",
"departments": [
"Financial Planning and Analysis"
],
"offices": [
"United States"
],
"url": "https://careers.airbnb.com/positions/8232207?gh_jid=8232207",
"updatedAt": "2026-09-29T20:00:54-04:00",
"firstPublished": "2026-09-25T14:53:46-04:00",
"requisitionId": "ONE",
"language": "en",
"descriptionText": "Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible fo…"
},
{
"jobId": 8184174,
"title": "Account Manager",
"companyName": "airbnb",
"location": "London, United Kingdom",
"departments": [
"Business Development"
],
"offices": [
"London, United Kingdom"
],
"url": "https://careers.airbnb.com/positions/8184174?gh_jid=8184174",
"updatedAt": "2026-09-29T20:00:54-04:00",
"firstPublished": "2026-09-09T04:35:19-04:00",
"requisitionId": "MULTI",
"language": "en",
"descriptionText": "Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible fo…"
},
{
"jobId": 7789554,
"title": "Associate Legal Counsel, Japan",
"companyName": "airbnb",
"location": "Tokyo, Japan",
"departments": [
"Legal"
],
"offices": [
"Tokyo, Japan"
],
"url": "https://careers.airbnb.com/positions/7789554?gh_jid=7789554",
"updatedAt": "2026-09-29T20:00:54-04:00",
"firstPublished": "2026-04-13T04:49:05-04:00",
"requisitionId": "ONE",
"language": "en",
"descriptionText": "Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible fo…"
}
]

Integrations and API

  • Apify API: start a run and get the results in one HTTP request:
curl -X POST "https://api.apify.com/v2/acts/compass_lab~greenhouse-jobs-scraper/run-sync-get-dataset-items?token=<YOUR_APIFY_TOKEN>" \
-H "Content-Type: application/json" -d '{"startUrls": [{"url": "https://job-boards.greenhouse.io/airbnb"}], "maxItems": 75, "includeDescription": true}'
  • Python (pip install apify-client):
from apify_client import ApifyClient
client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("compass_lab/greenhouse-jobs-scraper").call(run_input={"startUrls": [{"url": "https://job-boards.greenhouse.io/airbnb"}], "maxItems": 75, "includeDescription": true})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item)
  • JavaScript (npm install apify-client):
import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: '<YOUR_APIFY_TOKEN>' });
const run = await client.actor('compass_lab/greenhouse-jobs-scraper').call({"startUrls": [{"url": "https://job-boards.greenhouse.io/airbnb"}], "maxItems": 75, "includeDescription": true});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
  • Make, Zapier, n8n, Google Sheets, webhooks: use the Apify integrations (Integrations tab) to send each run's results where you need them, or to start a run from your workflow.
  • Schedules: run it hourly, daily or weekly from Apify Console (Schedules) and always have fresh jobs.

Tips and advanced options

  • Keep Max items and Max pages as low as you need: fewer pages means a faster, cheaper run.
  • Raise Delay between requests if the site responds slowly; keep Max concurrency low to stay polite.
  • Missing values are null. Fields that often come back empty are listed in the run log as data-quality warnings.
  • Companies: type company names (Stripe, Acme Inc) or paste job board URLs. Names are matched the way the platform writes them (Stripe Inc -> stripe); a company that can't be found is named in the log and skipped.
  • Daily monitoring: schedule the Actor with Only new jobs since the last run on. Each run returns only jobs it hasn't returned before for the same input. The IDs are kept in a named key-value store called <actor-name>-seen-<id> (the run log prints its name); delete that store in Storage > Key-value stores, or change the input, to start over.
  • Filters before charging: keywords, exclude keywords, location, remote only and posted after are applied before a job is saved, so filtered-out jobs cost nothing.

FAQ, disclaimers and support

Checked 2026-10-01. Greenhouse's Job Board API documentation states: "Job Board data is publicly available, so authentication is not required for any GET endpoints." robots.txt on boards-api.greenhouse.io and boards.greenhouse.io only disallows /embed/. Job postings are published by employers for public distribution; this actor collects job data, not personal data. Respect the API: low concurrency, 1 s between requests.

Our Actors are ethical and do not extract any private user data, such as email addresses, gender, or location. They only extract what the user has chosen to share publicly. We therefore believe that our Actors, when used for ethical purposes by Apify users, are safe. However, you should be aware that your results could contain personal data. Personal data is protected by the GDPR in the European Union and by other regulations around the world. You should not scrape personal data unless you have a legitimate reason to do so. If you're unsure whether your reason is legitimate, consult your lawyers.

How many results can I get?

Up to maxItems per run (0 means no limit), as many as the source lists. Each result is one dataset item, and you are only charged for items that are saved.

Can I run it on a schedule or from my own code?

Yes. Schedule it in Apify Console (Schedules), or call it with the run-sync-get-dataset-items endpoint or the Python/JavaScript clients shown in Integrations and API above.

What are the limitations?

  • Each board comes back in one response, so very large boards (thousands of jobs) make a longer run; use maxItems and the filters to keep it small.
  • Some companies leave fields empty (requisition ID, first published date, departments, offices); those come back as null or [].
  • A wrong or deactivated board token returns no jobs; the run log names the board that failed.
  • Greenhouse may change its API; the Actor checks every result and warns in the log if fields stop matching.
  • Job descriptions are written by each employer: use the data for search, monitoring and analysis, and link to the original posting rather than republishing descriptions in bulk.
  • descriptionText is capped at 5,000 characters.

Where can I get help?

Report problems or ideas on the Issues tab. To call this Actor from your own code, see the API tab.

Job Boards Suite: the same clean, typed output across sources, so you can combine them in one dataset.

ActorWhat it scrapesPrice
Breezy HR Jobs ScraperJob listings from Breezy HR$1.00 / 1,000
Lever Jobs ScraperJob listings from Lever$1.60 / 1,000
Python Jobs Scraper (python.org)Job listings from Python Jobs Scraper (python.org)$2.00 / 1,000
Recruitee Jobs ScraperJob listings from Recruitee$1.00 / 1,000
We Work Remotely Jobs ScraperJob listings from We Work Remotely$2.50 / 1,000
Working Nomads Jobs ScraperJob listings from Working Nomads$2.50 / 1,000