# a16z Scraper: Portfolio + Speedrun Companies (`getascraper/a16z-portfolio-scraper`) Actor

Scrape the full a16z portfolio (854 companies) and Speedrun accelerator (251 companies) in one run. Get social links, investment stage, exits, founder bios, and hiring signal. Export to Sheets, Excel, or the API.

- **URL**: https://apify.com/getascraper/a16z-portfolio-scraper.md
- **Developed by:** [GetAScraper](https://apify.com/getascraper) (community)
- **Categories:** Lead generation, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $3.75 / 1,000 company records

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## 🏢 a16z Scraper: Portfolio + Speedrun Companies

<table width="100%">
<tr>
<td style="padding:24px 28px;background:#FCF5F0;border:1px solid #F7E0D3;border-top:4px solid #D2540A;border-radius:12px">
<span style="font-size:23px;font-weight:800;color:#1C1917;line-height:1.3">Every a16z-backed company, and who's hiring right now</span><br>
<span style="font-size:15px;color:#57534E;line-height:1.6">The full a16z portfolio and the Speedrun accelerator roster in one run, with the hiring signal neither competing scraper tracks at all.</span>
</td>
</tr>
</table>

<table width="100%">
<tr>
<td style="padding:14px 12px;width:25%;background:#FFFFFF;border:1px solid #F7E0D3;border-radius:10px 0 0 10px;vertical-align:top">
<span style="font-size:15px;font-weight:800;color:#9A3E08">🏢 1,100+ companies, one run</span><br>
<span style="font-size:12px;color:#57534E">Full portfolio and Speedrun combined, not split across two actors.</span>
</td>
<td style="padding:14px 12px;width:25%;background:#FFFFFF;border:1px solid #F7E0D3;border-left:none;vertical-align:top">
<span style="font-size:15px;font-weight:800;color:#9A3E08">📞 Hiring signal, built in</span><br>
<span style="font-size:12px;color:#57534E">See which companies are hiring now, with a direct careers link and open role count.</span>
</td>
<td style="padding:14px 12px;width:25%;background:#FFFFFF;border:1px solid #F7E0D3;border-left:none;vertical-align:top">
<span style="font-size:15px;font-weight:800;color:#9A3E08">📈 Real exit and stage data</span><br>
<span style="font-size:12px;color:#57534E">Invest dates, exit dates, ticker symbols and acquirers for every public exit.</span>
</td>
<td style="padding:14px 12px;width:25%;background:#FFFFFF;border:1px solid #F7E0D3;border-left:none;border-radius:0 10px 10px 0;vertical-align:top">
<span style="font-size:15px;font-weight:800;color:#9A3E08">👥 Full founder bios</span><br>
<span style="font-size:12px;color:#57534E">Name, title, LinkedIn and bio for every Speedrun founder, not just a name.</span>
</td>
</tr>
</table>

Extract every company in the [a16z portfolio](https://a16z.com/portfolio/) and the [a16z Speedrun](https://speedrun.a16z.com/companies) accelerator. Export to JSON, CSV or Excel, or connect straight into Google Sheets and your own pipeline via the API. Run it on demand or on a schedule. No coding required.

### ✨ Why use this Actor

**Built for real workflows.**

- 💼 **Job seekers and recruiters**: filter by "backed by a16z" and know who's actually hiring, without checking five separate job boards by hand.
- 🎯 **Sales and BD teams**: use "funded by a16z" as an ICP filter for outreach, with a real website and LinkedIn page on every company, not a bare name.
- 📊 **VC analysts and researchers**: track a16z's exits and stage progression with real invest dates, exit dates, tickers, and acquirers.

**Both a16z directories, richer than either dedicated competitor.** One scraper covers only the Speedrun accelerator at $10 per 1,000 results. Another covers the full portfolio at $100 per 1,000 results, but with just 8 thin fields and no social links, no exit data, and no founder details. This Actor covers both in one run, with far more fields, at a fraction of the price.

### ⚙️ How it works

<table width="100%">
<tr>
<td style="padding:16px 14px;width:33%;background:#FCF5F0;border:1px solid #F7E0D3;border-radius:10px 0 0 10px;vertical-align:top">
<span style="font-size:12px;font-weight:800;color:#D2540A;letter-spacing:1px">STEP 1</span><br>
<span style="font-size:14px;font-weight:700;color:#1C1917">Pick your sources</span><br>
<span style="font-size:12px;color:#57534E">Choose portfolio, Speedrun, or both, then narrow with stage, cohort, industry or keywords.</span>
</td>
<td style="padding:16px 14px;width:33%;background:#FCF5F0;border:1px solid #F7E0D3;border-left:none;vertical-align:top">
<span style="font-size:12px;font-weight:800;color:#D2540A;letter-spacing:1px">STEP 2</span><br>
<span style="font-size:14px;font-weight:700;color:#1C1917">Run the Actor</span><br>
<span style="font-size:12px;color:#57534E">It pulls live from a16z's own sites and collects every matching company.</span>
</td>
<td style="padding:16px 14px;width:33%;background:#FCF5F0;border:1px solid #F7E0D3;border-left:none;border-radius:0 10px 10px 0;vertical-align:top">
<span style="font-size:12px;font-weight:800;color:#D2540A;letter-spacing:1px">STEP 3</span><br>
<span style="font-size:14px;font-weight:700;color:#1C1917">Get your data</span><br>
<span style="font-size:12px;color:#57534E">Download as JSON, CSV or Excel, or connect straight into Google Sheets or your own pipeline via the API.</span>
</td>
</tr>
</table>

### 📥 Input

| Field | Type | Required | Description |
|---|---|---|---|
| `sources` | enum | No | Which a16z directories to scrape: portfolio, Speedrun, or both. |
| `keywords` | string | No | Free-text match against company name, description, and tagline. |
| `stages` | enum | No | Portfolio only. Keep only companies at one of these investment stages. |
| `cohorts` | array | No | Speedrun only. Keep only companies from these cohorts, e.g. "SR003". |
| `industries` | array | No | Speedrun only. Keep only companies tagged with one of these industries. |
| `yearFoundedMin` / `yearFoundedMax` | integer | No | Only available for companies where a founding year is published. |
| `includeHiringSignal` | boolean | No | Check a16z's own jobs board and add whether each company is hiring, its careers link, and open role count. |
| `maxItems` | integer | No | Stop after collecting this many matching companies. |
| `proxyConfiguration` | object | No | Proxy settings. Datacenter proxy is enough. |

### 📤 Output

Every result is one row in the dataset. A typical portfolio item looks like this:

```json
{
  "source": "portfolio",
  "companyId": "373009",
  "name": "11x",
  "logoUrl": "https://d1lamhf6l6yk6d.cloudfront.net/uploads/2024/11/logo_001.svg",
  "websiteUrl": "https://www.11x.ai/",
  "linkedinUrl": "https://www.linkedin.com/company/11x-ai/",
  "xUrl": "https://twitter.com/11x_official",
  "stages": ["Venture"],
  "announcementExcerpt": "11x frees salespeople from the nightmare of administrative work.",
  "announcementUrl": "https://a16z.com/announcement/investing-in-11x/",
  "isHiring": true,
  "openRolesCount": 13,
  "careersUrl": "https://jobs.a16z.com/companies/11x.ai",
  "scrapedAt": "2026-08-29T04:19:39.625Z"
}
```

Download the dataset in JSON, CSV, Excel, HTML or XML from the Apify Console, or pull it through the API.

### 📊 Data table

| Field | Type | Description |
|---|---|---|
| `source` | string | `portfolio` or `speedrun`. |
| `companyId` / `name` | string | Unique ID and company name. |
| `logoUrl` / `websiteUrl` / `linkedinUrl` / `xUrl` | string | Logo and social links. |
| `description` / `foundedYear` | string/number | Overview and founding year, where published. |
| `stages` / `investDate` / `exitDate` / `tickerSymbol` / `acquirer` | array/string | Portfolio only. Investment stage history and exit details. |
| `announcementExcerpt` / `announcementUrl` | string | Portfolio only. a16z's own investment-thesis blog post. |
| `cohort` / `industries` / `preamble` / `keySignal` | string/array | Speedrun only. Program cohort and positioning. |
| `teamSize` / `city` / `state` / `country` / `region` | number/string | Speedrun only. Team size and location. |
| `founders` | array | Speedrun only. Name, title, bio, LinkedIn and photo per founder. |
| `isHiring` / `careersUrl` / `openRolesCount` | boolean/string/number | With hiring signal enabled. Whether the company is hiring, its careers page, and open role count. |

The Output tab also ships four pre-built views: all companies, portfolio stage and exits, Speedrun companies, and an unwound founders table.

### 💰 Pricing

This Actor is pay per result: you only pay for the companies you actually collect, and a run that returns nothing costs nothing. There is no subscription and no minimum spend.

### ⭐ Enjoying a16z Scraper?

<table width="100%">
<tr>
<td style="padding:20px 24px 14px;background:#FCF5F0;border:1px solid #F7E0D3;border-left:5px solid #D2540A;border-radius:10px 10px 0 0">
<span style="font-size:20px;letter-spacing:4px">⭐ ⭐ ⭐ ⭐ ⭐</span><br>
<span style="font-size:17px;font-weight:800;color:#1C1917">Saved you from checking five different job boards or scrapers?</span><br>
<span style="font-size:14px;color:#57534E">A 5-star rating takes 10 seconds and helps other founders, recruiters and analysts find it. Your feedback also tells us what to build next.</span>
</td>
</tr>
<tr>
<td style="padding:0;background:#D2540A;border:1px solid #F7E0D3;border-top:none;border-radius:0 0 10px 10px;text-align:center">
<a href="https://apify.com/getascraper/a16z-portfolio-scraper/reviews" style="display:block;padding:13px 16px;color:#FFFFFF;text-decoration:none;font-weight:800;font-size:15px;letter-spacing:0.3px">★&nbsp;&nbsp;Rate this Actor on Apify</a>
</td>
</tr>
</table>

### 🛠️ Tips for better runs

- Leave `includeHiringSignal` off for the fastest, cheapest runs when you only need company profiles.
- Turn `includeHiringSignal` on for job search or lead-gen use cases, it costs one extra request for the whole run, not per company.
- Set `maxItems` if you only need a sample. With both sources selected, portfolio companies are collected first, so a low `maxItems` may only return portfolio results; raise it or run each source separately for a balanced mix.

### ❓ FAQ

**Is it legal to scrape a16z's site?**
This Actor only collects data that is already publicly visible on a16z's own portfolio and Speedrun pages. You are responsible for how you use the data and for complying with a16z's terms of service.

**Why do some companies have missing fields?**
Not every field is published for every company. This Actor never invents or guesses a value: if a16z doesn't publish it, such as a founding year or invest date, the field is left out rather than filled with a placeholder.

**Does this include a "sector" or industry tag for portfolio companies?**
No. a16z's own portfolio data doesn't publish a sector field, so this Actor doesn't invent one. Speedrun companies do have a real `industries` field, which is included.

**Can I get notified of new investments or hires automatically?**
Yes. Schedule this Actor to run daily or weekly from the Apify Console and pipe new results into Google Sheets, Slack, or your own database with no code.

Found a bug or need a custom version of this Actor? Open an issue from the Actor's Issues tab and it'll be looked at directly.

### 🔗 Other actors

- [Startup investor contact database](https://apify.com/getascraper/startup-investor-contact-database) ↗ - publicly listed investor and firm contact data for fundraising outreach.
- [Ashby Jobs Scraper: Hiring Change Monitor](https://apify.com/getascraper/ashby-jobs-monitor) ↗ - tracks new and closed job postings on startup career pages.
- [Japan Company Scraper: 4.5M+ gBizINFO Records](https://apify.com/getascraper/gbizinfo-japan-company-scraper) ↗ - Japanese company registry data at scale.
- [US licensed contractor directory](https://apify.com/getascraper/us-licensed-contractor-directory) ↗ - licensed contractor business records by state.

# Actor input Schema

## `sources` (type: `array`):

Which a16z company directories to scrape.

## `keywords` (type: `string`):

Free-text match against company name, description, and tagline.

## `stages` (type: `array`):

Keep only portfolio companies at one of these stages.

## `cohorts` (type: `array`):

Keep only Speedrun companies from these cohorts, e.g. "SR003".

## `industries` (type: `array`):

Keep only Speedrun companies tagged with one of these industries, e.g. "Gaming", "AI".

## `yearFoundedMin` (type: `integer`):

Only available for companies where a founding year is published.

## `yearFoundedMax` (type: `integer`):

Only available for companies where a founding year is published.

## `includeHiringSignal` (type: `boolean`):

Check jobs.a16z.com for each company and add whether it's currently hiring, its careers page link, and open role count.

## `maxItems` (type: `integer`):

Stop after collecting this many matching companies.

## `proxyConfiguration` (type: `object`):

a16z's sites are fully open with no anti-bot protection; datacenter proxy is enough.

## Actor input object example

```json
{
  "sources": [
    "portfolio",
    "speedrun"
  ],
  "keywords": "",
  "stages": [],
  "cohorts": [],
  "industries": [],
  "includeHiringSignal": false,
  "maxItems": 1000,
  "proxyConfiguration": {
    "useApifyProxy": true
  }
}
```

# Actor output Schema

## `dataset` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "sources": [
        "portfolio",
        "speedrun"
    ],
    "keywords": "",
    "includeHiringSignal": false,
    "maxItems": 1000,
    "proxyConfiguration": {
        "useApifyProxy": true
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("getascraper/a16z-portfolio-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "sources": [
        "portfolio",
        "speedrun",
    ],
    "keywords": "",
    "includeHiringSignal": False,
    "maxItems": 1000,
    "proxyConfiguration": { "useApifyProxy": True },
}

# Run the Actor and wait for it to finish
run = client.actor("getascraper/a16z-portfolio-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "sources": [
    "portfolio",
    "speedrun"
  ],
  "keywords": "",
  "includeHiringSignal": false,
  "maxItems": 1000,
  "proxyConfiguration": {
    "useApifyProxy": true
  }
}' |
apify call getascraper/a16z-portfolio-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,getascraper/a16z-portfolio-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/5ZKc0NZ4pLc8CTtFK/builds/lnQwRLfZMETFKfEa5/openapi.json
