# Built In Jobs Scraper - Tech Jobs & Company Profiles (`scrapingmonkey/builtin-scraper`) Actor

Scrape Built In tech jobs and companies with skills, salary text, workplace labels and employer technology stacks for recruiting research. Choose the collection mode and set a result limit for each input.

- **URL**: https://apify.com/scrapingmonkey/builtin-scraper.md
- **Developed by:** [ScrapingMonkey](https://apify.com/scrapingmonkey) (community)
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

Collect public technology job listings and employer profiles with **Built In Scraper**. Combine skills and advertised pay from job cards with company technologies, industries and offices to research technology employers. Choose the records you need and export them to CSV, JSON or Excel without supplying a site login.

Use the results for US technology recruiting. Submit matching public inputs in one run, retain the original record links, and send the structured output to your recruitment spreadsheet, reporting dashboard or research database.

| At a glance | Details |
|---|---|
| 📥 Input | Search fields or links to the selected items; see the examples below |
| 📤 Output | One row per vacancy / company |
| 📄 Pagination | Automatic up to `maxResults` for each input, when more results are available |
| 🔐 Login required | No |
| ⚡ Processing | Multiple inputs per run with automatic retries for temporary errors |
| 💾 Delivery | One complete Apify dataset table; CSV, JSON and Excel export |

### What the Built In scraper collects 💼

Collect technology job listings and employer profiles. Combine skills and advertised pay from job cards with company technologies, industries and offices to research technology employers.

| Collection | What you receive |
|---|---|
| Built In jobs scraper | Records containing job titles, employers, locations, published pay, skills, ready for vacancy discovery and hiring-demand research. |
| Built In company profile scraper | Company descriptions, industry labels, office locations and published technology stacks. |
| Built In company directory scraper | Records containing company descriptions, locations, benefits, company size, ready for employer prospecting and recruitment account research. |

Data can include:

- Record identity: Record identifiers, public links and the published job title or company name.
- Pay and compensation: Published salary or wage text, plus the amount, range, currency and pay-period fields supplied with the record.
- Employer information: Employer names, identifiers and profile links, with the company information attached to the public record.
- Location and workplace: Published workplace and location information.
- Descriptions and requirements: Published job or company descriptions and listing summaries.
- Skills and qualifications: Required skills and experience or seniority requirements.
- Employment and benefits: Published benefits.
- Dates and hiring activity: Publication and update information.

### How to scrape Built In 🚀

1. Choose **Job Search** in **What to collect**, or select another collection mode.
2. Fill in the search fields, or paste links to the items you want to collect.
3. Set the maximum results and any optional filters.
4. Start the Actor and review the collected rows.
5. Export the dataset or connect it to your application.

#### Built In jobs scraper — input example

```json
{
  "mode": "job_search",
  "query": "software engineer",
  "maxResults": 10
}
```

`maxResults: 10` means up to 10 results for **each input**. It does not guarantee that the website has that many matches. You can set the limit as low as 1.

#### Built In company profile scraper — input example

```json
{
  "mode": "company_details",
  "companyUrls": [
    "https://builtin.com/company/jpmorgan-chase"
  ],
  "maxResults": 10
}
```

#### Built In company directory scraper — input example

```json
{
  "mode": "company_search",
  "maxResults": 10
}
```

### Built In data fields and complete output 📦

| Field group | Included data |
|---|---|
| Record identity | Record identifiers, public links and the published job title or company name. |
| Pay and compensation | Published salary or wage text, plus the amount, range, currency and pay-period fields supplied with the record. |
| Employer information | Employer names, identifiers and profile links, with the company information attached to the public record. |
| Location and workplace | Published workplace and location information. |
| Descriptions and requirements | Published job or company descriptions and listing summaries. |
| Skills and qualifications | Required skills and experience or seniority requirements. |
| Employment and benefits | Published benefits. |
| Dates and hiring activity | Publication and update information. |

Each collected item is a separate row. `input` identifies the submitted value and `result_type` identifies the selected mode. All returned fields appear in one table; columns that do not apply to a mode can be empty.

A complete JSON result is shown for every collection mode below, in the same order as the input examples. Each example includes every field for its mode; text and lists are not shortened. Values are representative and can change.

#### Built In jobs scraper — complete JSON result

```json
{
  "result_type": "job_search",
  "input": "https://builtin.com/jobs?country=USA&allLocations=true",
  "status": "success",
  "companyUrl": "https://builtin.com/company/wipfli",
  "locationText": "United States",
  "featured": false,
  "jobId": "6404562",
  "jobUrl": "https://builtin.com/job/tax-manager-manufacturing/6404562",
  "title": "Tax Manager - Manufacturing",
  "companyName": "Wipfli",
  "workplaceText": "Remote or Hybrid",
  "experienceLevel": "Senior level",
  "summary": "Manage tax compliance and advisory work, review tax returns, build client relationships, research tax matters, and mentor staff.",
  "skills": [
    "Accounting",
    "Cpa",
    "Tax Compliance"
  ],
  "salaryText": "106K-160K Annually",
  "postedText": "Reposted 40 Minutes Ago",
  "publishedAt": "2026-09-20T10:24:09"
}
```

#### Built In company profile scraper — complete JSON result

```json
{
  "result_type": "company_details",
  "input": "https://builtin.com/company/jpmorgan-chase",
  "status": "success",
  "companyId": "64026",
  "slug": "jpmorgan-chase",
  "name": "JPMorganChase",
  "profileUrl": "https://builtin.com/company/jpmorgan-chase",
  "tagline": "We’re one of the world’s biggest technology-driven companies",
  "descriptionText": "JPMorgan Chase & Co. (NYSE: JPM) is a leading global financial services firm with assets of $3.7 trillion and operations worldwide. The firm is a leader in investment banking, financial services for consumers and small businesses, commercial banking, financial transaction processing, and asset management. A component of the Dow Jones Industrial Average, JPMorgan Chase & Co. serves millions of consumers in the United States and many of the world’s most prominent corporate, institutional and government clients under its J.P. Morgan and Chase brands. Technology fuels every aspect of our company and is at the heart of everything we do. With over 50,000 technologists globally and an annual tech spend of $12 billion, we are dedicated to improving the design, analytics, development, coding, testing and application programming that goes into creating high quality software and new products. Learn more about technology at our firm, explore resources from our Distinguished Engineers, AI & ML researchers, and other experts; access the latest episode of our TechTrends podcast, and more at www.jpmorgan.com/technology. Information about JPMorgan Chase & Co. is available at www.jpmorganchase.com. ©2023 JPMorgan Chase & Co. All rights reserved. JPMorgan Chase is an Equal Opportunity Employer, including Disability/Veterans.",
  "industries": [
    "Financial Services"
  ],
  "offices.location": [
    "New York, New York, USA",
    "Bengaluru, Karnataka, IND",
    "Bournemouth, England",
    "Buenos Aires, Ciudad Autónoma de Buenos Aires, ARG",
    "Chicago, Illinois, USA",
    "Dallas, Texas, USA",
    "Dublin, Dublin, IRL",
    "Glasgow, Scotland",
    "Houston, Texas, USA",
    "Hyderabad, Telangana, IND",
    "London, England",
    "Mumbai, Maharashtra, IND",
    "New York, New York, USA",
    "Philadelphia, Pennsylvania, USA",
    "San Francisco, California, USA",
    "Singapore, Singapore, SGP",
    "Tampa, Florida, USA",
    "Westerville, Ohio, USA",
    "Wilmington, Delaware, USA"
  ],
  "offices.headquarters": [
    true,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false,
    false
  ],
  "technologies.name": [
    "Access",
    "ASP.NET",
    "C#",
    "C++",
    "Cassandra",
    "DB2",
    "Ember.js",
    "Hadoop",
    "Java",
    "JavaScript",
    "jQuery",
    "MariaDB",
    "Microsoft SQL Server",
    "MongoDB",
    "MySQL",
    "Node.js",
    "Oracle",
    "Python",
    "React",
    "Scala",
    "Spark",
    "Spring",
    "SQL",
    "Swift",
    "TensorFlow",
    "Dynatrace",
    "Splunk ",
    "Promtheus",
    "Grafana",
    "CloudFoundry",
    "Azure",
    "Kubernetes",
    "Confluence",
    "JIRA"
  ],
  "technologies.category": [
    "DATABASES",
    "FRAMEWORKS",
    "LANGUAGES",
    "LANGUAGES",
    "DATABASES",
    "DATABASES",
    "FRAMEWORKS",
    "FRAMEWORKS",
    "LANGUAGES",
    "LANGUAGES",
    "LIBRARIES",
    "DATABASES",
    "DATABASES",
    "DATABASES",
    "DATABASES",
    "FRAMEWORKS",
    "DATABASES",
    "LANGUAGES",
    "LIBRARIES",
    "LANGUAGES",
    "FRAMEWORKS",
    "FRAMEWORKS",
    "LANGUAGES",
    "LANGUAGES",
    "FRAMEWORKS",
    "DATABASES",
    "DATABASES",
    "DATABASES",
    "DATABASES",
    "LANGUAGES",
    "LANGUAGES",
    "LANGUAGES",
    "PROJECT MANAGEMENT",
    "PROJECT MANAGEMENT"
  ]
}
```

#### Built In company directory scraper — complete JSON result

```json
{
  "result_type": "company_search",
  "input": "https://builtin.com/companies",
  "status": "success",
  "companyId": "64510",
  "name": "RTB House",
  "descriptionText": "RTB House is a global leader in personalized advertising, using deep learning to deliver impactful, targeted campaigns. RTB House is a global technology company specializing in innovative marketing solutions powered by deep learning algorithms. Founded in 2012, the company has rapidly grown into a leader in the field of personalized advertising, offering a full-funnel marketing platform that drives real results. Our proprietary technology enables brands to deliver highly relevant and precisely targeted ads to consumers, enhancing engagement...",
  "companyUrl": "https://builtin.com/company/rtb-house",
  "logoUrl": "https://builtin.com/sites/www.builtin.com/files/2021-06/RTB House Logo.png",
  "industriesText": "AdTech • Artificial Intelligence • Big Data • Digital Media • eCommerce • Machine Learning • Marketing Tech",
  "locationText": "Fully Remote",
  "employeeCountText": "1,300 Employees",
  "benefitsCountText": "47 Benefits",
  "hiringNow": true,
  "featured": true
}
```

Images, tags and other lists stay with their parent result instead of creating extra rows. Invalid or unavailable inputs, inputs with no results, and requests that fail after all retries produce a row with `status: failed`. At most one failed row is saved per input; it includes `error_code` and `error_message` explaining the outcome. A normal end of pagination after collected results does not add a failed row.

### Input requirements and collection settings ⚙️

| Setting | What to enter |
|---|---|
| **What to collect** (`mode`) | Choose the information to collect. |
| **Maximum results per search or link** (`maxResults`) | Collect up to this many results for each search or link. Minimum: 1. |
| **Job title or keywords** (`query`) | Words to search for in job listings. Leave empty to include all jobs. For Job Search. |
| **Company page links** (`companyUrls`) | Paste one link per line. Example: https://builtin.com/company/jpmorgan-chase For Company Details. |

| Collection mode | Fields to use |
|---|---|
| Company Details | **Company page links** |
| Company Directory | Set the maximum results and start. |
| Job Search | **Job title or keywords** |

The result limit applies separately to each search or supplied link. Collection ends earlier when no more matching results are available.

### Built In scraping use cases 🎯

#### Hiring activity monitoring

Run repeat collections of Built In listings and compare vacancy titles and original links over time to identify newly advertised roles and changes in hiring activity.

#### Advertised compensation research

Compare the compensation published in Built In records. Keep the original salary text, units and available pay-period context when preparing recruitment or labour-market reports.

#### Employer qualification

Research organisations listed on Built In before adding them to recruitment or partnership lists. Retain their published identity, company description and available hiring context with the original profile link.

#### Skills demand analysis

Analyse the skills, experience or qualifications actually published in Built In records to compare hiring requirements across role groups and prepare recruitment briefs.

#### Working-condition comparison

Compare the contract, schedule or benefit information published in Built In records when researching how employers present their opportunities.

### Pricing and saved-result behavior 💰

See the Actor's **Pricing** tab for the active pricing model and current rate. Store settings may change, so this README does not claim a fixed cost.

Under dataset-item pricing:

- Each collected item saved in the dataset is one result.
- Images, tags and other values attached to that item do not create additional rows.
- Automatic retry attempts do not create extra dataset rows.
- A saved `failed` row is also one billable dataset item, including invalid inputs, unavailable items, no results, or exhausted retries. At most one failed row is saved per input, including when some results were already collected.
- `maxResults` caps all saved rows for each input, including its failed row. Repeated copies of the same input are processed once. More inputs or a higher limit can produce more saved results.

### Built In scraper API and integrations 🔌

Replace `$ACTOR_ID` with the ID shown in the Actor API tab and `$APIFY_TOKEN` with your token.

```bash
curl -X POST "https://api.apify.com/v2/acts/$ACTOR_ID/runs?token=$APIFY_TOKEN" \
  -H "Content-Type: application/json" \
  -d '{"mode":"job_search","query":"software engineer","maxResults":10}'
```

Use schedules for recurring collection, webhooks for completion notifications, and Apify integrations for Google Sheets, Make, Zapier, cloud storage, a CRM or your own application.

### Reliability, retries, and public-data limits ⚠️

Temporary connection errors, timeouts and access restrictions are retried automatically, up to five attempts. Invalid inputs save a failed row without an HTTP request; confirmed unavailable items save a failed row without unnecessary retries. Exhausted retries also save a failed row. Results already collected remain available if a later request fails, and normal pagination completion adds no error row. Collection also stops after two consecutive fully processed pages add no new results, even if the website provides another cursor. Required detail lookups are completed before checking a page; new results reset the counter, and independent reply branches are checked separately. Failures while initializing the Actor, saving data or scheduling requests stop the run instead of being reported as a failed website result.

The company directory is unfiltered. Job Search returns listing summaries; there is no separate full job-detail mode.

Built In controls which information is publicly available. Removed, private or restricted items may be unavailable, and optional values can be null or empty. A run can still fail if the results cannot be saved.

### Frequently asked questions ❓

#### Does it require an account on Built In?

No. You do not need to provide an account, password or login cookies. Only publicly available information is collected.

#### How does maxResults work?

It limits the number of saved results separately for every input. The minimum is 1. With maxResults set to 10, each input returns at most 10 available results. A detail lookup normally returns one record per input.

#### Which inputs and filters does this Actor support?

The company directory is unfiltered. Job Search returns listing summaries; there is no separate full job-detail mode.

#### What employer information can I collect?

Use the company mode to collect the public employer fields listed in the output section. The available profile data depends on what each organisation publishes on Built In.

#### Why are some fields empty?

The website may not publish that value, or the field may belong to a different collection mode. Missing values remain null or empty instead of being guessed.

#### Can I export the results to a spreadsheet?

Yes. Download CSV, Excel or JSON from the Apify dataset. All returned fields are included in the same results table.

### Support and responsible use 🛟

For a reproducible issue, share the run ID, selected mode, result limit and a safe public example through the support channel. Never disclose access tokens, passwords or proxy credentials.

Use public Built In information lawfully and responsibly. Follow applicable privacy, copyright, data-protection, contractual and platform requirements when storing, analyzing or redistributing results.

# Actor input Schema

## `mode` (type: `string`):

Choose the information to collect.

## `maxResults` (type: `integer`):

Collect up to this many results for each search or link. Minimum: 1.

## `query` (type: `string`):

Words to search for in job listings. Leave empty to include all jobs. For Job Search.

## `companyUrls` (type: `array`):

Paste one link per line. Example: https://builtin.com/company/jpmorgan-chase For Company Details.

## Actor input object example

```json
{
  "mode": "job_search",
  "maxResults": 10,
  "query": "software engineer"
}
```

# Actor output Schema

## `results` (type: `string`):

Collected entities and failed input outcomes, with all available fields and error explanations in one dataset.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "mode": "job_search",
    "maxResults": 10,
    "query": "software engineer"
};

// Run the Actor and wait for it to finish
const run = await client.actor("scrapingmonkey/builtin-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "mode": "job_search",
    "maxResults": 10,
    "query": "software engineer",
}

# Run the Actor and wait for it to finish
run = client.actor("scrapingmonkey/builtin-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "mode": "job_search",
  "maxResults": 10,
  "query": "software engineer"
}' |
apify call scrapingmonkey/builtin-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,scrapingmonkey/builtin-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/nMoGCce5XtPC91vk2/builds/w4sTgPmoadL0B8fjU/openapi.json
