# Owler Company Competitor Intel Scraper (`jungle_synthesizer/owler-company-competitor-intel-scraper`) Actor

Scrapes Owler company profiles for revenue estimates, employee counts, funding, CEO and firmographic data, plus the competitor graph and recent news headlines that make Owler distinctive as a company competitive-intelligence source.

- **URL**: https://apify.com/jungle\_synthesizer/owler-company-competitor-intel-scraper.md
- **Developed by:** [BowTiedRaccoon](https://apify.com/jungle_synthesizer) (community)
- **Categories:** Lead generation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.80 / 1,000 record scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Owler Company Revenue, Competitor & Funding Data Scraper

Scrape company profiles from [Owler](https://www.owler.com). Returns revenue estimates, employee counts, funding totals, CEO and headquarters details, plus the competitor graph and recent news headlines — the two things people actually go to Owler for.

***

### Owler Company Competitor Intel Scraper Features

- Extracts firmographic data: revenue estimate, employee count, headquarters, founding year, ownership, and industry
- Returns the full competitor graph per company — not just the top few names, but each competitor's own Owler URL
- Collects recent news, press, and blog headlines tied to the company
- Pulls CEO name straight off the profile, no LinkedIn detour required
- Look up specific companies by URL, or let it crawl broadly across Owler's company directory
- No account or API key required

***

### Who Uses Owler Company Data?

- **Sales teams** — build territory and account lists ranked by revenue and headcount, with the competitor set attached
- **Competitive intelligence analysts** — track a market's players and who competes with whom, without stitching it together by hand
- **Investors and analysts** — screen companies by funding raised and revenue band before digging into a filing
- **Market researchers** — map an industry's competitive landscape from a handful of seed companies
- **Recruiters** — cross-reference headcount and headquarters against a target list

***

### How Owler Company Competitor Intel Scraper Works

1. Give it specific company URLs, or leave the input blank
2. With no `companyUrls`, it walks Owler's company directory and pulls profiles up to your `maxItems` limit
3. Each profile page is parsed for firmographics, the competitor list, and recent news
4. Records land in your dataset as clean, flat JSON — competitor names and headlines included as plain strings, not nested objects you have to unpack

***

### Input

Default — crawl broadly:

```json
{
  "maxItems": 25
}
```

Look up specific companies:

```json
{
  "companyUrls": [
    "https://www.owler.com/company/openai",
    "https://www.owler.com/company/anthropic2"
  ],
  "maxItems": 10
}
```

| Field          | Type             | Default  | Description |
|----------------|------------------|----------|-------------|
| `maxItems`     | integer          | `10`     | Maximum number of company profiles to return. |
| `companyUrls`  | array of strings | *(none)* | Specific Owler company profile URLs to scrape (e.g. `https://www.owler.com/company/openai`). Leave empty to crawl broadly instead. |
| `resumeCursor` | string           | *(none)* | Continue a previous run without paying again for records you already received. See **Resuming a large crawl** below. |

#### Resuming a large crawl

Every run emits a `resumeCursor` in its Output. If a large crawl stops before it finishes — because it hit `maxItems`, your spend cap (`maxTotalChargeUsd`), or was aborted — start a new run with **the same input** plus that `resumeCursor` to continue from where it left off. The crawl resumes from the queued work the previous run didn't reach.

- You are **not re-charged** for records the earlier run already delivered.
- Resume within your account's run-retention window — on the free tier, roughly your 10 most recent runs. Once the source run is pruned, its `resumeCursor` is no longer valid.
- `resumeCursor` is opaque — supply it unmodified.

***

### Owler Company Competitor Intel Scraper Output Fields

```json
{
  "company_name": "OpenAI",
  "owler_url": "https://www.owler.com/company/openai",
  "estimated_annual_revenue": "$20B",
  "employees": 7800,
  "headquarters": "San Francisco, California",
  "founded_year": 2015,
  "ceo_name": "Sam Altman",
  "industry": "Software, Internet & Computer Services",
  "top_competitors": [
    "Anthropic — https://www.owler.com/company/anthropic2",
    "xAI — https://www.owler.com/company/xai",
    "Inflection AI — https://www.owler.com/company/inflectionai"
  ],
  "total_funding": "$298.9B",
  "website": "https://openai.com/",
  "description": "OpenAI is a California-based artificial intelligence research company that develops and deploys AI applications including the generative AI bot ChatGPT.",
  "recent_news_headlines": [
    "OpenAI and Anthropic probe tens of thousands of AI incidents, Axios reports — The Next Web",
    "OpenAI Unveils Sora: Text-to-Video AI Generates Realistic 1-Minute Clips — WebProNews"
  ]
}
```

| Field                      | Type             | Description                                                                           |
|----------------------------|------------------|---------------------------------------------------------------------------------------|
| `company_name`             | string           | Company display name                                                                  |
| `owler_url`                | string           | Canonical Owler company profile URL                                                   |
| `estimated_annual_revenue` | string           | Owler's estimated annual revenue, formatted (e.g. `$20B`)                             |
| `employees`                | integer          | Estimated employee count                                                              |
| `headquarters`             | string           | Headquarters city and state/region                                                    |
| `founded_year`             | integer          | Year the company was founded                                                          |
| `ceo_name`                 | string           | CEO (or top executive) display name                                                   |
| `industry`                 | string           | Industry / sector classification                                                      |
| `top_competitors`          | array of strings | Competitors as `"Name — Owler URL"` strings, drawn from Owler's full competitor graph |
| `total_funding`            | string           | Total funding raised, formatted (e.g. `$298.9B`)                                      |
| `website`                  | string           | Company's own website                                                                 |
| `description`              | string           | Company description                                                                   |
| `recent_news_headlines`    | array of strings | Recent news/press headlines as `"Title — Source"` strings                             |

***

### FAQ

#### Does this actor need an Owler account or API key?

No. Point it at a company and it returns the profile — no login, no API key.

#### Can I look up a single company instead of crawling broadly?

Yes. Pass one or more URLs in `companyUrls` and it scrapes exactly those profiles, nothing more.

#### How many competitors does each record include?

However many Owler lists on the profile — commonly a dozen or more, each with its own Owler URL so you can chain lookups.

#### What happens if a company profile has incomplete data?

Owler itself doesn't always have every field for every company (funding is a common gap for private companies with no disclosed rounds). Missing fields come back empty rather than guessed.

***

### Need More Features?

Need custom fields or a different target site? [File an issue](https://console.apify.com/actors/issues) or get in touch.

### Why Use Owler Company Competitor Intel Scraper?

- **Competitor graph included** — most company-data actors stop at firmographics. This one returns who each company competes with, and their Owler URLs, so you can chain lookups across a market.
- **Clean output** — flat JSON with consistent field names, competitor and headline arrays included as plain strings instead of objects you have to unpack.
- **Affordable** — pay per record, no subscription, no seat licenses.

# Actor input Schema

## `sp_intended_usage` (type: `string`):

What will this data feed? E.g. lead lists, KYB checks, price tracking.

## `sp_improvement_suggestions` (type: `string`):

Provide any feedback or suggestions for improvements.

## `sp_contact` (type: `string`):

We'll personally help with your use case. No spam.

## `resumeCursor` (type: `string`):

Leave empty for a fresh crawl. To CONTINUE a previous run where it stopped — without paying again for records you already received — paste the `resumeCursor` value from that run's Output (the run's OUTPUT key). Resume promptly: the previous run's data expires with your account's retention window (free tier: your ~10 most recent runs).

## `maxItems` (type: `integer`):

Maximum number of company profiles to scrape. Leave companyUrls empty to crawl broadly from Owler's company sitemap; set companyUrls to scrape specific companies instead.

## `companyUrls` (type: `array`):

Specific Owler company profile URLs to scrape (e.g. https://www.owler.com/company/openai). Leave empty to crawl broadly from Owler's company sitemap instead.

## Actor input object example

```json
{
  "sp_intended_usage": "Describe your intended use...",
  "sp_improvement_suggestions": "Share your suggestions here...",
  "sp_contact": "Share your email here...",
  "maxItems": 10
}
```

# Actor output Schema

## `results` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "sp_intended_usage": "Describe your intended use...",
    "sp_improvement_suggestions": "Share your suggestions here...",
    "sp_contact": "Share your email here...",
    "maxItems": 10
};

// Run the Actor and wait for it to finish
const run = await client.actor("jungle_synthesizer/owler-company-competitor-intel-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "sp_intended_usage": "Describe your intended use...",
    "sp_improvement_suggestions": "Share your suggestions here...",
    "sp_contact": "Share your email here...",
    "maxItems": 10,
}

# Run the Actor and wait for it to finish
run = client.actor("jungle_synthesizer/owler-company-competitor-intel-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "sp_intended_usage": "Describe your intended use...",
  "sp_improvement_suggestions": "Share your suggestions here...",
  "sp_contact": "Share your email here...",
  "maxItems": 10
}' |
apify call jungle_synthesizer/owler-company-competitor-intel-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,jungle_synthesizer/owler-company-competitor-intel-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/NCpiYBaJfLelGRJ8o/builds/Ie7GbWaPhlMlM8HPF/openapi.json
