# A16z Jobs Search Scraper (`alexist/a16z-jobs-search-scraper`) Actor

Scrape Andreessen Horowitz's job listings with precision. This scraper captures 38+ fields per posting—from job titles and apply URLs to salary, skills requirements, seniority levels, and company metadata—perfect for talent analytics, recruitment research, and career platform aggregation.

- **URL**: https://apify.com/alexist/a16z-jobs-search-scraper.md
- **Developed by:** [Alex](https://apify.com/alexist) (community)
- **Categories:** Automation, Developer tools, Jobs
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

Pay per usage

This Actor is paid per platform usage. The Actor is free to use, and you only pay for the Apify platform usage, which gets cheaper the higher subscription plan you have.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-usage

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## a16z Jobs Scraper: Extract Venture Capital Career Data

***

### What Is a16z.com?

Andreessen Horowitz (a16z) is a premier venture capital firm investing in software, crypto, biotech, and AI. Its careers portal—jobs.a16z.com—hosts openings across operations, engineering, finance, and strategy. These roles span early-stage startups to enterprise scale, attracting top talent globally. Manually extracting job details is tedious; the **a16z Jobs Scraper** automates this, unlocking structured data for recruitment intelligence, labor market research, and aggregator platforms.

***

### Overview

The **a16z Jobs Scraper** systematically extracts job listings from jobs.a16z.com's search results and detail pages, returning comprehensive records with 38+ fields per posting. It serves:

- **Recruiters** monitoring VC-backed hiring trends and competitor openings
- **Talent research teams** building candidate-skill correlation datasets
- **Career aggregators** syndicating a16z jobs to multi-platform boards
- **Market analysts** tracking venture hiring in tech and emerging sectors

The scraper is fast, reliable, and handles failures gracefully—perfect for large-scale collection or scheduled updates.

***

### Input Format

The scraper accepts a lightweight JSON configuration:

```json
{
  "urls": [
    "https://jobs.a16z.com/jobs"
  ],
  "ignore_url_failures": true,
  "max_items_per_url": 20
}
```

| Parameter | Type | Description |
|---|---|---|
| `urls` | Array | Job list URLs to scrape (e.g., search results or category pages). Start with `https://jobs.a16z.com/jobs` |
| `ignore_url_failures` | Boolean | If `true`, continues scraping even if individual URLs fail. Recommended for robust runs. |
| `max_items_per_url` | Integer | Maximum postings extracted per URL. Default: `20`. Increase for broader collection. |

> **Tip:** Use `ignore_url_failures: true` for bulk collection to avoid interruptions from temporary page issues.

***

### Output Format

**Sample output**

```json
{
  "apply_url": "https://careers.withwaymo.com/jobs?gh_jid=8068768",
  "company_domain": "waymo.com",
  "company_logos": {
    "manual": {
      "height": 160,
      "src": "https://dzh2zima160vx.cloudfront.net/logo/a04896dd1b05b675a7df4762df4fe370_112_160?Expires=1861920000&Signature=EdNw9DPJCYDLbjxU5eEwWJ1ELHjFa3WTlbbzZOXLlSmfSbgMtG32z6LdGGAtM8VmXGPbTMoIGR5WE~EPTSohV4ZZXDvdwEsBAvxFu-mta8yKoUHsWpruQPopFSBFK9k1XIrdKZLotubfhPpr~rZ6ggahzHJBjILvQ7PEHUefLMbocQlGRMeozW7oy1keH11J6qOjQyzJKmcps5s78-Jifvewo5DTlDnGRD0FAalL4mXh0Yl~V9LItmlv4w2aPn02zqNbqYCCeJy7qA6KQXAOAcWqCj2mHWibBuIwr8TIr61M~lni9s4x1gEuFJ~dg3EsZ~tkVuKmqHqYs-Bd3vsf7A__&Key-Pair-Id=APKAII5OVX4LZ3WT422Q",
      "width": 112
    },
    "linkedin": {
      "height": 100,
      "src": "https://dzh2zima160vx.cloudfront.net/profile/070485a5609a1e1bbf05d251e8bb13b5_100_100?Expires=1893456000&Signature=c755TEDKA4SHFcCs~zMMJXru0n~0voksSTJ0tjeMqtLff46ruNrFcYe7TGK6-4pMOpG-jWJCzyuOQ8vbG3cYGzmYjmnhIindTD2mu1TxByBniYJXfQsT17u6Sd5~uy6Di2Vfix3n67NM3AwwEPRe0dgJnUWzpIz4G6cXt7SnLKPez9DSgVRS0vfXYymd2yz31qj45rEjncwO2YQVWU0EHPHjfxxlDHijk4kD2DkxaTPtLypwmAQCG9fOXKbaAJQiEwzmFLuRIdKE2BB~squNkF39QmVa8D~NJpL5lSnNdjK7OveSMsRN62jcFieX7dCZgyOgrC3fXFi8BnsK3ISmvQ__&Key-Pair-Id=APKAII5OVX4LZ3WT422Q",
      "width": 100
    }
  },
  "company_id": "Waymo",
  "company_slug": "waymo",
  "company_name": "Waymo",
  "company_staff_count": 3661,
  "consider_hosted": false,
  "departments": [
    "Mfg (JDS)"
  ],
  "job_types": [
    {
      "id": "lead",
      "label": "Lead",
      "value": "lead"
    },
    {
      "id": "procurement",
      "label": "Procurement Manager",
      "value": "procurement"
    }
  ],
  "job_functions": [
    {
      "id": "Supply Chain",
      "label": "Supply Chain",
      "value": "Supply Chain"
    }
  ],
  "locations": [
    "Mountain View, CA, USA"
  ],
  "normalized_locations": [
    {
      "id": "Mountain View, California",
      "label": "Mountain View, California",
      "value": "Mountain View, California"
    }
  ],
  "salary": {
    "period": {
      "label": "Year",
      "value": "year"
    },
    "min_value": 200000,
    "max_value": 247000,
    "currency": {
      "label": "USD",
      "value": "USD"
    },
    "is_original": true
  },
  "skills": [
    {
      "id": "resume:Project Management",
      "label": "Project Management",
      "value": "resume:Project Management"
    },
    {
      "id": "resume:SAP",
      "label": "SAP",
      "value": "resume:SAP"
    },
    {
      "id": "resume:Strategic Thinking",
      "label": "Strategic Thinking",
      "value": "resume:Strategic Thinking"
    },
    {
      "id": "resume:Management Consulting",
      "label": "Management Consulting",
      "value": "resume:Management Consulting"
    },
    {
      "id": "resume:Procurement",
      "label": "Procurement",
      "value": "resume:Procurement"
    },
    {
      "id": "resume:English",
      "label": "English",
      "value": "resume:English"
    }
  ],
  "required_skills": [
    {
      "id": "resume:Project Management",
      "label": "Project Management",
      "value": "resume:Project Management"
    },
    {
      "id": "resume:SAP",
      "label": "SAP",
      "value": "resume:SAP"
    },
    {
      "id": "resume:Strategic Thinking",
      "label": "Strategic Thinking",
      "value": "resume:Strategic Thinking"
    }
  ],
  "preferred_skills": [
    {
      "id": "resume:Management Consulting",
      "label": "Management Consulting",
      "value": "resume:Management Consulting"
    },
    {
      "id": "resume:Procurement",
      "label": "Procurement",
      "value": "resume:Procurement"
    },
    {
      "id": "resume:English",
      "label": "English",
      "value": "resume:English"
    }
  ],
  "manager": false,
  "consultant": false,
  "contractor": false,
  "consider_levels": [
    [
      5,
      6
    ],
    [
      7,
      8
    ]
  ],
  "min_years_exp": 10,
  "regions": [
    {
      "id": "North America",
      "label": "North America",
      "value": "North America"
    },
    {
      "id": "Pacific States",
      "label": "Pacific States",
      "value": "Pacific States"
    },
    {
      "id": "West Coast",
      "label": "West Coast",
      "value": "West Coast"
    }
  ],
  "stages": [
    {
      "id": "1000+ employees",
      "label": "1000+ employees",
      "value": "1000+ employees"
    },
    {
      "id": "Growth",
      "label": "Growth",
      "value": "Growth"
    }
  ],
  "funding_lv": null,
  "markets": [
    {
      "id": "Enterprise",
      "label": "Enterprise",
      "value": "Enterprise"
    }
  ],
  "time_stamp": "2026-07-16T18:21:43Z",
  "title": "Aftermarket Planning & Procurement Lead",
  "url": "https://careers.withwaymo.com/jobs?gh_jid=8068768",
  "job_id": "8068768",
  "remote": false,
  "hybrid": false,
  "scores": {
    "score": 1001.1924151334345,
    "match_score": 0,
    "age_score": 0.9924151334344136,
    "richness_score": 2
  },
  "ats_jobs": [],
  "matching_talent": {
    "count": 0,
    "matches": []
  },
  "job_seniorities": [
    {
      "id": "mid",
      "label": "Mid",
      "value": "mid"
    },
    {
      "id": "senior",
      "label": "Senior",
      "value": "senior"
    }
  ],
  "job_seniority_ids": [
    "mid",
    "senior"
  ],
  "is_featured": false,
  "from_url": "https://jobs.a16z.com/jobs"
}
```

Each job record includes 38 fields across multiple dimensions:

#### Apply & URL Data

| Field | Description |
|---|---|
| `Apply URL` | Direct link to the application form or external careers page |
| `URL` | Canonical URL of the job listing |
| `Job ID` | Unique identifier for the posting in a16z's system |

#### Company Information

| Field | Description |
|---|---|
| `Company Name` | Name of the hiring company or portfolio firm |
| `Company Domain` | Company website domain |
| `Company Logos` | Logo image URLs for branding |
| `Company ID` | Internal company identifier |
| `Company Slug` | URL-friendly company name |
| `Company Staff Count` | Headcount or size category |
| `Funding Level` | VC funding stage (Seed, Series A, B, C, etc.) |
| `Stages` | Growth stage classification |
| `Markets` | Industry vertical or market focus |

#### Job Details

| Field | Description |
|---|---|
| `Title` | Job title (e.g., "Senior Software Engineer") |
| `Departments` | Org unit (e.g., Engineering, Operations, Finance) |
| `Job Functions` | Functional role (e.g., Product Management, Sales) |
| `Job Types` | Employment type (Full-time, Part-time, Contract) |
| `Time Stamp` | When the listing was posted or last updated |
| `Is Featured` | Whether the job is highlighted or promoted |
| `ATS Jobs` | Applicant Tracking System integration status |

#### Location & Remote Work

| Field | Description |
|---|---|
| `Locations` | City and country of the role |
| `Normalized Locations` | Standardized location format for analysis |
| `Regions` | Geographic region classification |
| `Remote` | Fully remote position flag |
| `Hybrid` | Hybrid work arrangement flag |

#### Compensation & Experience

| Field | Description |
|---|---|
| `Salary` | Salary range or compensation band |
| `Min Years Exp` | Minimum years of experience required |
| `Consider Levels` | Career level (Junior, Mid, Senior, Lead) |
| `Job Seniorities` | Seniority tier(s) the role targets |
| `Job Seniority IDs` | Internal seniority level identifiers |

#### Skills & Requirements

| Field | Description |
|---|---|
| `Skills` | All skills associated with the role |
| `Required Skills` | Must-have technical or soft skills |
| `Preferred Skills` | Nice-to-have qualifications |

#### Hiring Model & Talent

| Field | Description |
|---|---|
| `Manager` | Whether the role includes management responsibility |
| `Consultant` | Consultant or advisory role flag |
| `Contractor` | Contract or freelance availability flag |
| `Consider Hosted` | Candidate placement or hosted talent option |
| `Matching Talent` | Count of matched candidate profiles |
| `Scores` | Job quality or match scoring metrics |

***

### How to Use

1. **Identify target URLs** — Visit jobs.a16z.com/jobs or filtered search results. Copy the URL(s).
2. **Configure inputs** — Paste URLs into the `urls` array. Set `max_items_per_url` based on expected listing count.
3. **Enable resilience** — Set `ignore_url_failures: true` to skip temporary failures without stopping.
4. **Execute** — Run the scraper and monitor the activity log.
5. **Download & integrate** — Export as JSON or CSV for analysis, dashboards, or databases.

**Common tips:**

- Collect during off-peak hours to minimize server load.
- Start with `max_items_per_url: 50` for a broad snapshot.
- Filter URLs by department (Engineering, Operations) to narrow results.

***

### Use Cases & Business Value

- **Hiring trends analysis:** Track which roles, skills, and seniorities a16z portfolio companies prioritize
- **Salary benchmarking:** Collect compensation data to inform offer strategies
- **Skill demand mapping:** Identify emerging skills and experience requirements across tech sectors
- **Recruitment outreach:** Feed structured job data into multi-channel aggregators or talent platforms
- **Market intelligence:** Monitor portfolio company growth by new headcount and hiring velocity

The scraper delivers market-grade data suitable for dashboards, ML models, and strategic decision-making.

***

### Conclusion

The **a16z Jobs Scraper** provides venture-grade access to a16z's job postings at scale. With 38 richly detailed fields per listing, it streamlines recruitment research, talent analytics, and competitive intelligence. Deploy it today to unlock venture hiring insights.

# Actor input Schema

## `urls` (type: `array`):

Add the URLs of the Jobs list urls you want to scrape. You can paste URLs one by one, or use the Bulk edit section to add a prepared list.

## `ignore_url_failures` (type: `boolean`):

If true, the scraper will continue running even if some URLs fail to be scraped.

## `max_items_per_url` (type: `integer`):

The maximum number of items to scrape per URL.

## Actor input object example

```json
{
  "urls": [
    "https://jobs.a16z.com/jobs"
  ],
  "ignore_url_failures": true,
  "max_items_per_url": 20
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "urls": [
        "https://jobs.a16z.com/jobs"
    ],
    "ignore_url_failures": true,
    "max_items_per_url": 20
};

// Run the Actor and wait for it to finish
const run = await client.actor("alexist/a16z-jobs-search-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "urls": ["https://jobs.a16z.com/jobs"],
    "ignore_url_failures": True,
    "max_items_per_url": 20,
}

# Run the Actor and wait for it to finish
run = client.actor("alexist/a16z-jobs-search-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "urls": [
    "https://jobs.a16z.com/jobs"
  ],
  "ignore_url_failures": true,
  "max_items_per_url": 20
}' |
apify call alexist/a16z-jobs-search-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=alexist/a16z-jobs-search-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/acts/JnKpdYHYliZgFjYjI/builds/1RhCSbLQF2IHWdmfZ/openapi.json
