Canada Job Bank Jobs Scraper avatar

Canada Job Bank Jobs Scraper

Pricing

$1.00 / 1,000 jobs

Go to Apify Store
Canada Job Bank Jobs Scraper

Canada Job Bank Jobs Scraper

Canada Job Bank jobs scraper: export job postings from Canada's Job Bank open data (title, NOC occupation, province, city, employment type, salary range, posting date) to JSON, CSV, Excel. Official open data (OGL-Canada), no login, keyword and location filters. Track new, changed and closed jobs.

Pricing

$1.00 / 1,000 jobs

Rating

0.0

(0)

Developer

COMPASSLAB

COMPASSLAB

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Categories

Share

Get Canada Job Bank job listings as clean JSON, CSV or Excel, from the Apify API or on a schedule. No login, about $1.00 for 1,000 jobs.

What does Canada Job Bank Jobs Scraper do?

Canada Job Bank Jobs Scraper extracts structured data from open.canada.ca. Extract Job Bank records published as open data on open.canada.ca, with dataset and resource metadata attached to each row, so job market analysis and monitoring don't require manual CSV downloads. It works as an API for open.canada.ca data: run it from Apify Console, on a schedule, or from your own code, and get clean, typed JSON with numbers as numbers and dates in ISO 8601.

What you get

Data19 fields per item: postingId, jobTitle, nocCode, nocName, firstPostingDate, vacancyCount, ...
FormatsJSON, CSV, Excel, HTML, or the Apify API
Price$1.00 per 1,000 jobs, pay per result
AccessPublic open.canada.ca data only: no login, no cookies, robots.txt respected
LicenceOpen Government Licence – Canada

Why use Canada Job Bank Jobs Scraper?

  • Recruiters and staffing agencies: see which companies are hiring, for which roles and where, and reach out first.
  • Job aggregators and job boards: feed fresh, de-duplicated listings into your own board on a schedule.
  • Market research and HR analytics: track hiring trends, salaries (where published), locations and remote share.
  • AI agents and RAG: give an LLM live, structured job data instead of stale web pages.

Main features:

  • Follows pagination up to maxPages pages per start URL and stops at maxItems results.
  • Filters: keywords (keywords), excludeKeywords (exclude keywords), location (location), remoteOnly (remote only), postedAfter (posted after), resourceFormats (resource formats), keyword (keyword filter), onlyChanges (only changes since the last run), onlyNewSinceLastRun (only new jobs since the last run), pageSize (page size), so you only get (and pay for) the results you need.
  • Polite by default: respects robots.txt, at most maxConcurrency parallel requests and a delay between requests.
  • Checks every result against field validators, so layout changes show up as clear data-quality warnings.
  • Runs on the Apify platform: scheduling, API access, integrations, monitoring and datasets you can export.

What data can Canada Job Bank Jobs Scraper extract?

FieldTypeDescription
postingIdstringJob Bank posting (location snapshot) ID
jobTitlestringJob title
nocCodestringNOC 2021 occupation code
nocNamestringNOC 2021 occupation name
firstPostingDateISO 8601 dateFirst posting date (ISO 8601)
vacancyCountintegerNumber of vacancies
provincestringProvince or territory
citystringCity (null when not given)
employmentTypestringFull time / part time
employmentTermstringe.g. Permanent employment
salaryMinnumberMinimum salary
salaryMaxnumberMaximum salary
salaryPerstringSalary period: Hour, Year, ...
salaryDetailstringSalary as written in the posting
teleworkbooleanTelework offered (null when not stated)
monthstringMonthly file the posting comes from
urlstringCKAN dataset API URL the record was read from
changeTypestringnew, changed, unchanged, closed since the last run with the same input
changesobjectWhat changed since the last run: {field: {from, to}}

How to scrape open.canada.ca

  1. Open Canada Job Bank Jobs Scraper in Apify Console and go to the Input tab.
  2. Enter what to scrape (see the Input section below), for example the start URLs.
  3. Set Max items to the number of results you need.
  4. Click Start and wait for the run to finish.
  5. Download the results from the Output tab, or fetch them with the API.

How much will it cost to scrape open.canada.ca?

This Actor is priced per result: $1.00 per 1,000 results, with no extra charge for platform usage. That is about $1.00 for 1,000 jobs: 100 results cost $0.10 and 10,000 results cost $10.00. Set a maximum cost per run and the Actor stops when it is reached. Filters (companies, keywords, location, remote only, posted after, only new jobs) run before charging: you never pay for jobs you filtered out.

Input

See the Input tab for full configuration options.

FieldTypeRequiredDescription
keywordsarraynoKeep jobs whose title (and description, when 'Include description' is on) contains any of these words. Case-insensitive.
excludeKeywordsarraynoDrop jobs whose title (and description, when on) contains any of these words.
locationstringnoKeep jobs whose location contains this text, e.g. 'Berlin' or 'United States'.
remoteOnlybooleannoKeep only remote jobs (location says Remote/Anywhere, or the source marks the job as remote).
postedAfterstringnoKeep jobs posted on or after this date (YYYY-MM-DD). Jobs without a date are kept.
resourceFormatsarraynoOnly read dataset resources whose format is in this list.
keywordstringnoOnly return rows where any column contains this text (case-insensitive). Leave empty for all rows.
onlyChangesbooleannoFor scheduled runs: return (and pay for) only new, changed and closed records, compared with the last run with the same input. The snapshot is kept in a named key-value store; delete it to start over.
onlyNewSinceLastRunbooleannoFor scheduled runs: skip jobs this same input already returned. The IDs are kept in a named key-value store; delete it to start over.
maxItemsintegernoMaximum number of items to return (0 = unlimited).
startUrlsarrayyesopen.canada.ca package_show API URLs (or dataset page URLs containing the dataset id) to read.
pageSizeintegernoRows requested per page when paging through a resource (limit/offset).
maxPagesintegernoMaximum listing pages to follow per start URL (pagination).
maxConcurrencyintegernoMaximum parallel requests (politeness; 1-10).
requestDelayMsintegernoMinimum delay between requests, in milliseconds (at least 250).
proxyTypestringnonone (direct connection), datacenter (Apify Proxy, cheapest) or residential (opt-in, billed per GB, fewer blocks). The actor never switches by itself.
proxyCountrystringnoTwo-letter country code for the proxy IP (optional).

Example input:

{
"startUrls": [
{
"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"
}
],
"resourceFormats": [
"CSV",
"JSON"
],
"pageSize": 100,
"maxItems": 150,
"maxPages": 3,
"maxConcurrency": 2,
"requestDelayMs": 1000,
"proxyType": "none",
"keyword": ""
}

Output

You can download the dataset in various formats such as JSON, HTML, CSV, or Excel. Example results from a real run:

[
{
"postingId": "15774648",
"jobTitle": "heavy equipment operator",
"nocCode": "73400",
"nocName": "Heavy equipment operators",
"firstPostingDate": "2026-08-07",
"vacancyCount": 1,
"province": "Alberta",
"city": null,
"employmentType": "Full time",
"employmentTerm": "Permanent employment",
"salaryMin": 33.77,
"salaryMax": 35.85,
"salaryPer": "Hour",
"salaryDetail": "$33.77 to $35.85 hourly",
"telework": null,
"month": "August 2026 Job Postings Advertised on Canada's National Job Bank Website",
"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"
},
{
"postingId": "15798624",
"jobTitle": "truck trailer mechanic",
"nocCode": "72410",
"nocName": "Automotive service technicians, truck and bus mechanics and mechanical repairers",
"firstPostingDate": "2026-08-17",
"vacancyCount": 1,
"province": "Alberta",
"city": null,
"employmentType": "Full time",
"employmentTerm": "Permanent employment",
"salaryMin": 38,
"salaryMax": 40,
"salaryPer": "Hour",
"salaryDetail": "$38.00 to $40.00 hourly",
"telework": null,
"month": "August 2026 Job Postings Advertised on Canada's National Job Bank Website",
"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"
},
{
"postingId": "15798969",
"jobTitle": "recreation leader",
"nocCode": "54100",
"nocName": "Program leaders and instructors in recreation, sport and fitness",
"firstPostingDate": "2026-08-17",
"vacancyCount": 1,
"province": "Alberta",
"city": null,
"employmentType": "Part time",
"employmentTerm": "Term or contract",
"salaryMin": 21.46,
"salaryMax": 23.81,
"salaryPer": "Hour",
"salaryDetail": "$21.46 to $23.81 hourly",
"telework": null,
"month": "August 2026 Job Postings Advertised on Canada's National Job Bank Website",
"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"
}
]

Integrations and API

  • Apify API: start a run and get the results in one HTTP request:
curl -X POST "https://api.apify.com/v2/acts/compass_lab~canada-jobbank-scraper/run-sync-get-dataset-items?token=<YOUR_APIFY_TOKEN>" \
-H "Content-Type: application/json" -d '{"startUrls": [{"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"}], "resourceFormats": ["CSV", "JSON"], "pageSize": 100, "maxItems": 150, "keyword": ""}'
  • Python (pip install apify-client):
from apify_client import ApifyClient
client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("compass_lab/canada-jobbank-scraper").call(run_input={"startUrls": [{"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"}], "resourceFormats": ["CSV", "JSON"], "pageSize": 100, "maxItems": 150, "keyword": ""})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
print(item)
  • JavaScript (npm install apify-client):
import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: '<YOUR_APIFY_TOKEN>' });
const run = await client.actor('compass_lab/canada-jobbank-scraper').call({"startUrls": [{"url": "https://open.canada.ca/data/api/action/package_show?id=ea639e28-c0fc-48bf-b5dd-b8899bd43072"}], "resourceFormats": ["CSV", "JSON"], "pageSize": 100, "maxItems": 150, "keyword": ""});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
  • Make, Zapier, n8n, Google Sheets, webhooks: use the Apify integrations (Integrations tab) to send each run's results where you need them, or to start a run from your workflow.
  • Schedules: run it hourly, daily or weekly from Apify Console (Schedules) and always have fresh jobs.

Tips and advanced options

  • Keep Max items and Max pages as low as you need: fewer pages means a faster, cheaper run.
  • Raise Delay between requests if the site responds slowly; keep Max concurrency low to stay polite.
  • Missing values are null. Fields that often come back empty are listed in the run log as data-quality warnings.
  • Companies: type company names (Stripe, Acme Inc) or paste job board URLs. Names are matched the way the platform writes them (Stripe Inc -> stripe); a company that can't be found is named in the log and skipped.
  • Daily monitoring: schedule the Actor with Only new jobs since the last run on. Each run returns only jobs it hasn't returned before for the same input. The IDs are kept in a named key-value store called <actor-name>-seen-<id> (the run log prints its name); delete that store in Storage > Key-value stores, or change the input, to start over.
  • Filters before charging: keywords, exclude keywords, location, remote only and posted after are applied before a job is saved, so filtered-out jobs cost nothing.

Track changes

Schedule the Actor (for example every morning) with Only changes since the last run on. Each run then returns, and charges for, only the jobs that moved since the previous run with the same input:

  • new: a job this input hadn't returned before;
  • changed: same job, but its title, location, salary, type or department changed; changes says what, e.g. {"salaryMax": {"from": 90000, "to": 100000}};
  • closed: a job returned last time that is gone now (charged like a job). Closed jobs are reported only when the run read the whole board: raise Max pages if a board has more pages than it allows.

Without that option every job is returned as usual, each marked new, changed or unchanged.

Add it to your usual input, e.g. "onlyChanges": true.

The last snapshot is kept in a named key-value store called <actor-name>-watch-<id> (the run log prints its name). To start over, delete that store in Storage > Key-value stores, or change the input. Only new jobs since the last run still works as before (new jobs only, its own store).

FAQ, disclaimers and support

Vetted on 2026-10-02: Green tier (autonomy policy, 2026-10-02): robots.txt exists and permits /data/api/action/package_show (no matching Disallow); public endpoint; platform family 'open-data-ckan' confirmed by Mouad on 2026-10-01: Open-source portal software, government open-data licences

The data is published under the Open Government Licence – Canada: you may reuse it, including commercially, provided you credit the source as the licence requires.

Our Actors are ethical and do not extract any private user data, such as email addresses, gender, or location. They only extract what the user has chosen to share publicly. We therefore believe that our Actors, when used for ethical purposes by Apify users, are safe. However, you should be aware that your results could contain personal data. Personal data is protected by the GDPR in the European Union and by other regulations around the world. You should not scrape personal data unless you have a legitimate reason to do so. If you're unsure whether your reason is legitimate, consult your lawyers.

How many results can I get?

Up to maxItems per run (0 means no limit), as many as the source lists. Each result is one dataset item, and you are only charged for items that are saved.

Can I run it on a schedule or from my own code?

Yes. Schedule it in Apify Console (Schedules), or call it with the run-sync-get-dataset-items endpoint or the Python/JavaScript clients shown in Integrations and API above.

What are the limitations?

  • Only the columns published by the Government of Canada are available. The data is open-data extracts, not a live feed of the Job Bank website, so it may lag behind current postings.
  • Column names and file layouts can change when the publisher updates the dataset, which changes the keys inside the record field.
  • Very large resource files make long runs. Use maxPages, maxItems or the keyword filter to keep runs short and cheap.
  • Files may be published in English and French, so text values can appear in either language.
  • Dataset ids or resource links can be retired or moved by the publisher, in which case that start URL returns no rows.
  • Data is used under the Open Government Licence - Canada. Check the licence terms before republishing.

Where can I get help?

Report problems or ideas on the Issues tab. To call this Actor from your own code, see the API tab.

Job Boards Suite: the same clean, typed output across sources, so you can combine them in one dataset.

ActorWhat it scrapesPrice
Breezy HR Jobs ScraperJob listings from Breezy HR$1.00 / 1,000
Greenhouse Jobs ScraperJob listings from Greenhouse$1.60 / 1,000
Lever Jobs ScraperJob listings from Lever$1.60 / 1,000
Pinpoint Jobs ScraperJob listings from Pinpoint$1.00 / 1,000
Python Jobs Scraper (python.org)Job listings from Python Jobs Scraper (python.org)$2.00 / 1,000
Recruitee Jobs ScraperJob listings from Recruitee$1.00 / 1,000
We Work Remotely Jobs ScraperJob listings from We Work Remotely$2.50 / 1,000
Working Nomads Jobs ScraperJob listings from Working Nomads$2.50 / 1,000