# Workday Jobs Scraper — Any Company Career Site (`haketa/workday-scraper`) Actor

Scrape jobs from any company's Workday careers site (myworkdayjobs.com): title, location, job ID, posted date, remote type, full description, time type and dates. Paste any Workday careers URL.

- **URL**: https://apify.com/haketa/workday-scraper.md
- **Developed by:** [Haketa](https://apify.com/haketa) (community)
- **Categories:** Jobs
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Workday Jobs Scraper — Any Company's Career Site

> **Extract every job from any company's Workday careers site: title, location, job ID, posted date, remote type, full description, time type and dates.** Thousands of large employers run their careers on Workday (myworkdayjobs.com) — paste a company's Workday URL and get clean, structured JSON/CSV/Excel in seconds. Built for recruiters, job boards, sourcers and market researchers.

[![Workday](https://img.shields.io/badge/Source-Workday%20Career%20Sites-005cb9)]()
[![Any Company](https://img.shields.io/badge/Any-Company%20Tenant-1aa06d)]()
[![Full Details](https://img.shields.io/badge/Includes-Full%20Descriptions-8250df)]()
[![Export](https://img.shields.io/badge/Export-JSON%20%2F%20CSV%20%2F%20Excel-fb8500)]()

***

### What This Actor Does

**Workday** powers the careers sites of thousands of large enterprises (Fortune 500s and beyond). Each company has its own Workday careers site at `{company}.wd{N}.myworkdayjobs.com`. This Actor reads jobs straight from Workday's own public data feed and returns each opening as a clean row:

- **Job basics** — title, company, job ID, location, posted date, remote type
- **Full details (optional)** — complete job description, exact primary location, additional locations, time type (full/part time), start date, country
- **Links** — the direct public job-posting URL

Point it at one or many Workday career sites and it paginates through the entire job list.

***

### Why Use This

- **Go straight to the source.** Company career sites have jobs that aggregators miss or list late — and no third-party noise. This is the employer's own live feed.
- **Any Workday company.** Just paste the careers URL — it works for any tenant on myworkdayjobs.com, no per-company setup.
- **Complete and structured.** Title, location, dates, remote type and full descriptions as clean fields — ready to search, filter and load.
- **Fast and cheap.** Reads Workday's public JSON API directly — no headless browser — so it stays quick and inexpensive across thousands of jobs.

***

### Quick Start

#### Run it in the console (no code)

1. Open a company's Workday careers page (e.g. search "NVIDIA careers Workday").
2. Copy the URL — it looks like `https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite`.
3. Paste it into **Workday careers URLs**, optionally set a **Search keyword** and toggle **Include full job details**.
4. Click **Start**, then export as **JSON, CSV, Excel or HTML**, or push to Google Sheets, a webhook or a database.

#### Run it via API (Python)

```python
from apify_client import ApifyClient

client = ApifyClient("YOUR_APIFY_TOKEN")

run_input = {
    "startUrls": ["https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite"],
    "searchText": "engineer",
    "includeDescription": True,
    "maxItems": 500,
}

run = client.actor("YOUR_USERNAME/workday-scraper").call(run_input=run_input)

for job in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(job["title"], "·", job["location"], "·", job["timeType"])
```

#### Track several companies at once (Python)

```python
run = client.actor("YOUR_USERNAME/workday-scraper").call(run_input={
    "startUrls": [
        "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite",
        "https://redhat.wd5.myworkdayjobs.com/jobs",
    ],
    "maxItems": 2000,
})
rows = list(client.dataset(run["defaultDatasetId"]).iterate_items())
by_company = {}
for r in rows:
    by_company[r["company"]] = by_company.get(r["company"], 0) + 1
print(by_company)  # open roles per company
```

#### Find remote roles (Node.js)

```javascript
import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: 'YOUR_APIFY_TOKEN' });

const run = await client.actor('YOUR_USERNAME/workday-scraper').call({
    startUrls: ['https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite'],
    includeDescription: true,
    maxItems: 1000,
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
const remote = items.filter(j => /remote/i.test(j.remoteType) || /remote/i.test(j.location));
console.log(remote.length, 'remote roles');
```

***

### Input Parameters

| Field | Type | Description |
|---|---|---|
| `startUrls` | array | Workday careers URLs (`https://{company}.wd{N}.myworkdayjobs.com/{site}`). One or many companies. |
| `searchText` | string | Optional keyword filter (e.g. `engineer`, `marketing`). Empty = all jobs. |
| `includeDescription` | boolean | Fetch each job's full description, exact locations, time type and dates. Default `false`. |
| `maxItems` | integer | Max jobs across all sites. `0` = no limit (paginate to the end). |
| `proxyConfiguration` | object | Apify Proxy. Datacenter is enough and enabled by default. |

**Finding a Workday URL:** on a company's site, click "Careers" / "Jobs" — if the address bar shows `…myworkdayjobs.com/…`, copy it. That's all you need.

***

### Output

Each job is one record (with `includeDescription: true`):

```json
{
  "company": "nvidia",
  "jobId": "JR2020549",
  "title": "Senior Software Engineer, PyTorch - Deep Learning",
  "location": "6 Locations",
  "locationDetail": "US, CA, Santa Clara",
  "additionalLocations": ["US, WA, Remote", "US, MA, Remote"],
  "country": "United States of America",
  "remoteType": "",
  "timeType": "Full time",
  "postedOn": "Posted Yesterday",
  "startDate": "2026-09-24",
  "description": "We are now looking for a Senior Deep Learning Software Engineer, PyTorch…",
  "url": "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/..._JR2020549",
  "scrapedAt": "2026-09-25T11:34:51.000Z"
}
```

Without `includeDescription`, each record still has company, title, location, posted date, remote type, job ID and URL.

***

### Use Cases

#### 1. Recruiting & competitive intelligence

Track exactly which roles a target company is hiring for, where and when — org growth, team build-outs and hiring shifts, straight from their own careers feed.

#### 2. Job boards & aggregation

Pull complete, up-to-the-minute job lists from employer career sites to power a niche job board or feed — no reliance on slow aggregators.

#### 3. Talent sourcing & market mapping

Map open roles by function, location and remote type across companies to time outreach and understand talent demand.

#### 4. Labor-market & sector research

Quantify hiring volume and role mix across a set of employers over time — a clean signal for analysts and investors.

#### 5. Sales & lead timing

A surge in a company's hiring (e.g. sales, RevOps, engineering) is a buying signal — monitor it to time B2B outreach.

#### 6. HR & compensation benchmarking

Compare role titles, seniority, locations and time types across competitors for workforce planning.

***

### Tips

- **Any tenant works:** the URL pattern is `{company}.wd{N}.myworkdayjobs.com/{site}` (the `wd{N}` and site name vary by company — just copy the whole URL).
- **`searchText`** filters server-side, so it's efficient for large tenants (e.g. only `data engineer` roles).
- **`includeDescription`** adds one request per job for full text — enable it when you need descriptions.
- **`maxItems: 0`** paginates to the end; set a cap for quick samples.
- **Schedule it** with Apify Schedules to keep a fresh, deduped job feed per company.

***

### Frequently Asked Questions

**Which companies does this support?**
Any employer whose careers site runs on Workday (URL contains `myworkdayjobs.com`) — thousands of large companies.

**Do I need a login?**
No. The Actor reads publicly visible job postings — no account required.

**How do I get a company's Workday URL?**
Open the company's careers page and copy the address if it contains `myworkdayjobs.com`. Paste the whole URL.

**Can I scrape multiple companies in one run?**
Yes — add several Workday URLs to `startUrls`; each job row carries its `company`.

**What export formats are supported?**
JSON, CSV, Excel, HTML, or via API — plus Google Sheets, webhooks, Make and Zapier.

***

### Legal & Responsible Use

This Actor collects only publicly available job-posting information for research, recruiting and analytics. You are responsible for how you use the data. Please:

- Respect each site's Terms of Service and robots directives.
- Comply with applicable data-protection laws when handling any personal data.
- Do not use the data for spam, harassment, or any unlawful purpose.
- Use reasonable request volumes and scheduling.

This project is an independent tool and is not affiliated with, endorsed by, or sponsored by Workday or any employer.

# Actor input Schema

## `startUrls` (type: `array`):

Company Workday careers URLs, in the form https://{company}.wd{N}.myworkdayjobs.com/{site} (e.g. https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite). Open a company's Workday careers page and copy the URL.

## `searchText` (type: `string`):

Optional keyword to filter jobs (e.g. 'engineer', 'marketing'). Leave empty for all jobs.

## `includeDescription` (type: `boolean`):

Fetch each job's full description, exact location, additional locations, time type and dates. Slower (one extra request per job) but much richer.

## `maxItems` (type: `integer`):

Maximum number of jobs across all career sites. 0 = no limit (paginate to the end).

## `proxyConfiguration` (type: `object`):

Apify Proxy. Datacenter is enough for Workday and is enabled by default; add residential only if you hit rate limits.

## Actor input object example

```json
{
  "startUrls": [
    "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite"
  ],
  "includeDescription": true,
  "maxItems": 25,
  "proxyConfiguration": {
    "useApifyProxy": true
  }
}
```

# Actor output Schema

## `company` (type: `string`):

Workday tenant (company)

## `jobId` (type: `string`):

Job requisition ID

## `title` (type: `string`):

Job title

## `location` (type: `string`):

Location (list text)

## `locationDetail` (type: `string`):

Exact primary location

## `country` (type: `string`):

Country

## `remoteType` (type: `string`):

Remote / hybrid / on-site

## `timeType` (type: `string`):

Full time / Part time

## `postedOn` (type: `string`):

Posted date text

## `startDate` (type: `string`):

Posting start date

## `description` (type: `string`):

Full job description (text)

## `externalPath` (type: `string`):

Workday job path

## `url` (type: `string`):

Job posting URL

## `scrapedAt` (type: `string`):

ISO timestamp

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite"
    ],
    "includeDescription": true,
    "maxItems": 25,
    "proxyConfiguration": {
        "useApifyProxy": true
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("haketa/workday-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "startUrls": ["https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite"],
    "includeDescription": True,
    "maxItems": 25,
    "proxyConfiguration": { "useApifyProxy": True },
}

# Run the Actor and wait for it to finish
run = client.actor("haketa/workday-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    "https://nvidia.wd5.myworkdayjobs.com/NVIDIAExternalCareerSite"
  ],
  "includeDescription": true,
  "maxItems": 25,
  "proxyConfiguration": {
    "useApifyProxy": true
  }
}' |
apify call haketa/workday-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,haketa/workday-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/VVH1xrVWUk1XUIcBT/builds/UczIvoSAIZwlbxOUS/openapi.json
