# We Work Remotely Jobs Scraper (`compass_lab/weworkremotely-jobs-scraper`) Actor

We Work Remotely jobs scraper: export remote job listings from weworkremotely.com (title, company, region, category, company HQ, date, description) to JSON, CSV or Excel. Uses the site's public RSS feeds, no login, keyword filter.

- **URL**: https://apify.com/compass\_lab/weworkremotely-jobs-scraper.md
- **Developed by:** [COMPASSLAB](https://apify.com/compass_lab) (community)
- **Categories:** Jobs, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

Pay per event

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

**Get We Work Remotely job listings as clean JSON, CSV or Excel, from the Apify API or on a schedule. No login, about $2.50 for 1,000 jobs, $6.30 with full descriptions.**

### What does We Work Remotely Jobs Scraper do?

**We Work Remotely Jobs Scraper** extracts structured data from **[weworkremotely.com](https://weworkremotely.com)**. Collect remote job listings from We Work Remotely's public RSS feeds, with an optional title keyword filter, for job boards, lead generation and market research. It works
as an **API for weworkremotely.com data**: run it from Apify Console, on a schedule, or from your own code, and get clean,
typed JSON with numbers as numbers and dates in ISO 8601.

### What you get

| | |
|---|---|
| **Data** | 10 fields per item: `url`, `title`, `company`, `region`, `category`, `companyHeadquarters`, ... |
| **Formats** | JSON, CSV, Excel, HTML, or the Apify API |
| **Price** | $2.50 per 1,000 jobs, $6.30 with full descriptions, pay per result |
| **Access** | Public weworkremotely.com data only: no login, no cookies, robots.txt respected |

### Why use We Work Remotely Jobs Scraper?

- **Recruiters and staffing agencies**: see which companies are hiring, for which roles and where, and reach out first.
- **Job aggregators and job boards**: feed fresh, de-duplicated listings into your own board on a schedule.
- **Market research and HR analytics**: track hiring trends, salaries (where published), locations and remote share.
- **AI agents and RAG**: give an LLM live, structured job data instead of stale web pages.

Main features:

- Gets each source's results in as few requests as possible and stops at `maxItems` results.
- Filters: `keywords` (keywords), `excludeKeywords` (exclude keywords), `location` (location), `remoteOnly` (remote only), `postedAfter` (posted after), `keyword` (title keyword), `includeDescription` (include description), `onlyNewSinceLastRun` (only new jobs since the last run), so you only get (and pay for) the results you need.
- Polite by default: respects robots.txt, at most `maxConcurrency` parallel requests and a delay between requests.
- Checks every result against field validators, so layout changes show up as clear data-quality warnings.
- Runs on the Apify platform: scheduling, API access, integrations, monitoring and datasets you can export.

### What data can We Work Remotely Jobs Scraper extract?

| Field | Type | Description |
|---|---|---|
| `url` | string | Job posting URL (the item's <link>) |
| `title` | string | Job title (the part after 'Company: ' in the RSS item title) |
| `company` | string | Hiring company (the part before the first ': ' in the RSS item title) |
| `region` | string | Allowed work region from <region>, e.g. 'Anywhere in the World' or 'USA Only' |
| `category` | string | Job category from <category>, e.g. Programming |
| `companyHeadquarters` | string | Company headquarters as given in the ad's 'Headquarters:' line (e.g. 'New York, NY' or 'Remote') |
| `companyUrl` | string | Company website from the ad's 'URL:' line |
| `postedAt` | ISO 8601 date | Publication time in ISO 8601, converted from the RFC 2822 <pubDate> |
| `descriptionText` | string | Job description with HTML stripped, truncated to 5000 characters |
| `feedUrl` | string | The RSS feed URL the job was read from |

### How to scrape weworkremotely.com

1. Open We Work Remotely Jobs Scraper in Apify Console and go to the Input tab.
2. Enter what to scrape (see the Input section below), for example the start URLs.
3. Set **Max items** to the number of results you need.
4. Click **Start** and wait for the run to finish.
5. Download the results from the Output tab, or fetch them with the API.

### How much will it cost to scrape weworkremotely.com?

This Actor is priced **per result**: $2.50 per 1,000 results, with no extra charge for platform usage. That is about $2.50 for 1,000 jobs, $6.30 with full descriptions: 100 results cost $0.25 and 10,000 results cost $25.00. Set a maximum cost per run and the Actor stops when it is reached. With **Include description** on, each job comes with its full description text and is priced **$6.30 per 1,000 results** (job with description). Filters (companies, keywords, location, remote only, posted after, only new jobs) run **before** charging: you never pay for jobs you filtered out.

### Input

See the Input tab for full configuration options.

| Field | Type | Required | Description |
|---|---|---|---|
| `keywords` | array | no | Keep jobs whose title (and description, when 'Include description' is on) contains any of these words. Case-insensitive. |
| `excludeKeywords` | array | no | Drop jobs whose title (and description, when on) contains any of these words. |
| `location` | string | no | Keep jobs whose location contains this text, e.g. 'Berlin' or 'United States'. |
| `remoteOnly` | boolean | no | Keep only remote jobs (location says Remote/Anywhere, or the source marks the job as remote). |
| `postedAfter` | string | no | Keep jobs posted on or after this date (YYYY-MM-DD). Jobs without a date are kept. |
| `keyword` | string | no | Optional. Only return jobs whose title contains this text (case-insensitive). |
| `includeDescription` | boolean | no | Add the full job description text. Priced as 'job with description'. |
| `onlyNewSinceLastRun` | boolean | no | For scheduled runs: skip jobs this same input already returned. The IDs are kept in a named key-value store; delete it to start over. |
| `maxItems` | integer | no | Maximum number of items to return (0 = unlimited). |
| `startUrls` | array | no | We Work Remotely RSS feed URLs, e.g. https://weworkremotely.com/categories/remote-programming-jobs.rss, https://weworkremotely.com/categories/remote-design-jobs.rss or https://weworkremotely.com/remote-jobs.rss. |
| `maxPages` | integer | no | Maximum listing pages to follow per start URL (pagination). |
| `maxConcurrency` | integer | no | Maximum parallel requests (politeness; 1-10). |
| `requestDelayMs` | integer | no | Minimum delay between requests, in milliseconds (at least 250). |
| `proxyType` | string | no | none (direct connection), datacenter (Apify Proxy, cheapest) or residential (opt-in, billed per GB, fewer blocks). The actor never switches by itself. |
| `proxyCountry` | string | no | Two-letter country code for the proxy IP (optional). |

Example input:

```json
{
  "startUrls": [
    {
      "url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
    }
  ],
  "maxItems": 100,
  "maxPages": 3,
  "maxConcurrency": 2,
  "requestDelayMs": 1000,
  "proxyType": "none",
  "includeDescription": true
}
```

### Output

You can download the dataset in various formats such as JSON, HTML, CSV, or Excel. Example results
from a real run:

```json
[
  {
    "url": "https://weworkremotely.com/remote-jobs/samsara-staff-software-engineer",
    "title": "Staff Software Engineer",
    "company": "Samsara",
    "region": "Anywhere in the World",
    "category": "Full-Stack Programming",
    "companyHeadquarters": "Remote",
    "companyUrl": "https://www.samsara.com/",
    "postedAt": "2026-08-17T13:57:14+00:00",
    "descriptionText": "About the role: Samsara (NYSE: IOT) sits at the center of hardware, software, AI, and the physical world. The platform processes 25+ trillion data points annually from IoT sensors, cameras, and connected devices across thousands of organizations, and that dataset is compounding: it grew from 4 trill…",
    "feedUrl": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
  },
  {
    "url": "https://weworkremotely.com/remote-jobs/sticker-mule-ai-agent-engineer",
    "title": "AI agent engineer",
    "company": "Sticker Mule",
    "region": "Anywhere in the World",
    "category": "Full-Stack Programming",
    "companyHeadquarters": "New York, NY",
    "companyUrl": "https://www.stickermule.com",
    "postedAt": "2026-09-16T12:24:17+00:00",
    "descriptionText": "Sticker Mule is building the most lucrative commerce platform on the Internet by combining software, manufacturing, and AI in one stack. We're hiring an engineer to build, run and manage a team of AI agents to help us innovate faster, improve performance, and better serve customers. Work performed I…",
    "feedUrl": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
  },
  {
    "url": "https://weworkremotely.com/remote-jobs/acculynx-senior-software-engineer",
    "title": "Senior Software Engineer",
    "company": "AccuLynx",
    "region": "Anywhere in the World",
    "category": "Full-Stack Programming",
    "companyHeadquarters": "Beloit, WI",
    "companyUrl": "http://acculynx.com",
    "postedAt": "2026-09-15T14:12:47+00:00",
    "descriptionText": "Senior Software Engineer AccuLynx is a fast-growing SaaS provider of business management software for roofing contractors. With more than 10 years in the business and impressive year-over-year revenue growth, we have quickly established ourselves as the leading software product in this multi-billion…",
    "feedUrl": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
  }
]
```

### Integrations and API

- **Apify API**: start a run and get the results in one HTTP request:

```bash
curl -X POST "https://api.apify.com/v2/acts/compass_lab~weworkremotely-jobs-scraper/run-sync-get-dataset-items?token=<YOUR_APIFY_TOKEN>" \
  -H "Content-Type: application/json" -d '{"startUrls": [{"url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"}], "maxItems": 100, "includeDescription": true}'
```

- **Python** (`pip install apify-client`):

```python
from apify_client import ApifyClient

client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("compass_lab/weworkremotely-jobs-scraper").call(run_input={"startUrls": [{"url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"}], "maxItems": 100, "includeDescription": true})
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)
```

- **JavaScript** (`npm install apify-client`):

```javascript
import { ApifyClient } from 'apify-client';

const client = new ApifyClient({ token: '<YOUR_APIFY_TOKEN>' });
const run = await client.actor('compass_lab/weworkremotely-jobs-scraper').call({"startUrls": [{"url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"}], "maxItems": 100, "includeDescription": true});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
```

- **Make, Zapier, n8n, Google Sheets, webhooks**: use the Apify integrations (Integrations tab) to send each
  run's results where you need them, or to start a run from your workflow.
- **Schedules**: run it hourly, daily or weekly from Apify Console (Schedules) and always have fresh jobs.

### Tips and advanced options

- Keep **Max items** and **Max pages** as low as you need: fewer pages means a faster, cheaper run.
- Raise **Delay between requests** if the site responds slowly; keep **Max concurrency** low to stay polite.
- Missing values are `null`. Fields that often come back empty are listed in the run log as data-quality warnings.
- **Companies**: type company names (`Stripe`, `Acme Inc`) or paste job board URLs. Names are matched the way
  the platform writes them (`Stripe Inc` -> `stripe`); a company that can't be found is named in the log and skipped.
- **Daily monitoring**: schedule the Actor with **Only new jobs since the last run** on. Each run returns only jobs
  it hasn't returned before for the same input. The IDs are kept in a named key-value store called
  `<actor-name>-seen-<id>` (the run log prints its name); delete that store in Storage > Key-value stores, or change
  the input, to start over.
- **Filters before charging**: keywords, exclude keywords, location, remote only and posted after are applied before
  a job is saved, so filtered-out jobs cost nothing.

### FAQ, disclaimers and support

#### Is it legal to scrape weworkremotely.com?

Checked 2026-10-01. We Work Remotely publishes public RSS feeds (linked in the site footer); its Terms and Conditions contain no clause against automated access or scraping. robots.txt allows everything except account/admin pages; the actor only reads the RSS feeds. Job data only, no personal data. Link back to the job URL when you republish.

Our Actors are ethical and do not extract any private user data, such as email addresses, gender, or location.
They only extract what the user has chosen to share publicly. We therefore believe that our Actors, when used
for ethical purposes by Apify users, are safe. However, you should be aware that your results could contain
personal data. Personal data is protected by the GDPR in the European Union and by other regulations around the
world. You should not scrape personal data unless you have a legitimate reason to do so. If you're unsure whether
your reason is legitimate, consult your lawyers.

#### How many results can I get?

Up to `maxItems` per run (0 means no limit), as many as the source lists. Each result is one dataset item, and
you are only charged for items that are saved.

#### Can I run it on a schedule or from my own code?

Yes. Schedule it in Apify Console (Schedules), or call it with the `run-sync-get-dataset-items` endpoint or the
Python/JavaScript clients shown in **Integrations and API** above.

#### What are the limitations?

- Each RSS feed holds the latest ~25 jobs of its category; add more category feeds to `startUrls` for more results, and schedule runs to build a history.
- Company headquarters and website come from the ad text and are missing for some ads (`null`).
- Some promoted ads stay in the feed for a long time; use `postedAt` to filter old ones.
- Job content belongs to the posting companies and We Work Remotely: link to the original ad when you show it.

#### Where can I get help?

Report problems or ideas on the Issues tab. To call this Actor from your own code, see the API tab.

### Related actors

**Job Boards Suite**: the same clean, typed output across sources, so you can combine them in one dataset.

| Actor | What it scrapes | Price |
|---|---|---|
| [Breezy HR Jobs Scraper](https://apify.com/compass_lab/breezy-jobs-scraper) | Job listings from Breezy HR | $1.00 / 1,000 |
| [Greenhouse Jobs Scraper](https://apify.com/compass_lab/greenhouse-jobs-scraper) | Job listings from Greenhouse | $1.60 / 1,000 |
| [Lever Jobs Scraper](https://apify.com/compass_lab/lever-jobs-scraper) | Job listings from Lever | $1.60 / 1,000 |
| [Python Jobs Scraper (python.org)](https://apify.com/compass_lab/python-job-board-scraper) | Job listings from Python Jobs Scraper (python.org) | $2.00 / 1,000 |
| [Recruitee Jobs Scraper](https://apify.com/compass_lab/recruitee-jobs-scraper) | Job listings from Recruitee | $1.00 / 1,000 |
| [Working Nomads Jobs Scraper](https://apify.com/compass_lab/workingnomads-jobs-scraper) | Job listings from Working Nomads | $2.50 / 1,000 |

# Actor input Schema

## `keywords` (type: `array`):

Keep jobs whose title (and description, when 'Include description' is on) contains any of these words. Case-insensitive.

## `excludeKeywords` (type: `array`):

Drop jobs whose title (and description, when on) contains any of these words.

## `location` (type: `string`):

Keep jobs whose location contains this text, e.g. 'Berlin' or 'United States'.

## `remoteOnly` (type: `boolean`):

Keep only remote jobs (location says Remote/Anywhere, or the source marks the job as remote).

## `postedAfter` (type: `string`):

Keep jobs posted on or after this date (YYYY-MM-DD). Jobs without a date are kept.

## `includeDescription` (type: `boolean`):

Add the full job description text. Changes the price: $6.30 per 1,000 jobs instead of $2.50.

## `onlyNewSinceLastRun` (type: `boolean`):

For scheduled runs: skip jobs this same input already returned. The IDs are kept in a named key-value store; delete it to start over.

## `maxItems` (type: `integer`):

Maximum number of items to return (0 = unlimited). You pay per result, so this also caps the cost.

## `startUrls` (type: `array`):

We Work Remotely RSS feed URLs, e.g. https://weworkremotely.com/categories/remote-programming-jobs.rss, https://weworkremotely.com/categories/remote-design-jobs.rss or https://weworkremotely.com/remote-jobs.rss.

## `maxPages` (type: `integer`):

Maximum listing pages to follow per start URL (pagination).

## `maxConcurrency` (type: `integer`):

Maximum parallel requests (politeness; 1-10).

## `requestDelayMs` (type: `integer`):

Minimum delay between requests, in milliseconds (at least 250).

## `proxyType` (type: `string`):

none (direct connection), datacenter (Apify Proxy, cheapest) or residential (opt-in, billed per GB, fewer blocks). The actor never switches by itself.

## `proxyCountry` (type: `string`):

Two-letter country code for the proxy IP (optional).

## Actor input object example

```json
{
  "remoteOnly": false,
  "includeDescription": false,
  "onlyNewSinceLastRun": false,
  "maxItems": 20,
  "startUrls": [
    {
      "url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
    }
  ],
  "maxPages": 3,
  "maxConcurrency": 2,
  "requestDelayMs": 1000,
  "proxyType": "none"
}
```

# Actor output Schema

## `dataset` (type: `string`):

All scraped items (overview view)

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "includeDescription": false,
    "maxItems": 20,
    "startUrls": [
        {
            "url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
        }
    ],
    "maxPages": 3,
    "maxConcurrency": 2,
    "requestDelayMs": 1000,
    "proxyType": "none"
};

// Run the Actor and wait for it to finish
const run = await client.actor("compass_lab/weworkremotely-jobs-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "includeDescription": False,
    "maxItems": 20,
    "startUrls": [{ "url": "https://weworkremotely.com/categories/remote-programming-jobs.rss" }],
    "maxPages": 3,
    "maxConcurrency": 2,
    "requestDelayMs": 1000,
    "proxyType": "none",
}

# Run the Actor and wait for it to finish
run = client.actor("compass_lab/weworkremotely-jobs-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "includeDescription": false,
  "maxItems": 20,
  "startUrls": [
    {
      "url": "https://weworkremotely.com/categories/remote-programming-jobs.rss"
    }
  ],
  "maxPages": 3,
  "maxConcurrency": 2,
  "requestDelayMs": 1000,
  "proxyType": "none"
}' |
apify call compass_lab/weworkremotely-jobs-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,compass_lab/weworkremotely-jobs-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/e7mDSZeVFHarrRMn0/builds/Zq1e8EInDnbxl35EP/openapi.json
