# Company Enrichment from Domain (Logo, Socials, Tech) (`forevertools/company-website-enrichment`) Actor

Company data enrichment / lead enrichment from domains: clean company info & social links — name, description, logo, social profiles (LinkedIn/X/Facebook/Instagram/YouTube/GitHub), address, tech stack, email provider, domain age. Firmographic company data only — no personal emails.

- **URL**: https://apify.com/forevertools/company-website-enrichment.md
- **Developed by:** [Forever Tools](https://apify.com/forevertools) (community)
- **Categories:** Lead generation, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$2.00 / 1,000 company enricheds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Company Enrichment from Domain — Logo, Social Profiles, Tech Stack, Domain Age

Turn a list of **company domains** into clean **company profiles**, read live from each company's own website:
name, description, logo, favicon, official social pages (LinkedIn, X/Twitter, Facebook, Instagram, YouTube, TikTok,
GitHub, Pinterest, Crunchbase), address & founding date from structured data, **tech stack** (130+ technologies),
**email provider** (Google Workspace, Microsoft 365, …) and **domain age**.

Great for: CRM enrichment and cleanup, lead scoring (e.g. "Shopify stores on Microsoft 365"), filling logos in
your app, market maps, and AI agents that need quick facts about a company.

**Company data only.** The actor does not collect personal email addresses, employee names or personal profiles.

### Enrich a list of company domains in bulk

Give the actor domains or URLs (`stripe.com`, `https://www.notion.so`). For each one it loads the homepage once, reads
the page metadata and the schema.org Organization data (JSON-LD) if the site publishes it, and returns one row per
company. Optional DNS and RDAP lookups add the email provider and the domain's registration date. Because the data is
read from the company's own site, it reflects what the company publishes about itself today, not a third-party database.

### What you get per company

- **Identity:** `name`, `legalName`, `title`, `description`, `language`, `url`, `statusCode`.
- **Branding:** `logo`, `ogImage`, `favicon`.
- **Social profiles:** `linkedin`, `twitter`, `facebook`, `instagram`, `youtube`, `tiktok`, `github`, `pinterest`, `crunchbase` (company pages only, not personal profiles or share links).
- **Location and contact from structured data:** `foundingDate`, `phone`, `streetAddress`, `city`, `region`, `postalCode`, `country`. Only present when the site publishes schema.org data.
- **Tech stack:** `technologyNames` and `technologyCategories`, detected from headers, cookies, scripts, meta tags and HTML.
- **Domain info:** `emailProvider` (from MX records), `domainCreatedAt`, `domainAgeYears`.
- **Meta:** `input`, `domain`, `enrichedAt`, and `warning` or `error` when something went wrong.

### Find a company's logo, social links and tech stack from its domain

Typical jobs: fetch a logo and favicon to display next to a company in your app; collect the LinkedIn and X pages of
a list of accounts; or see which CMS, ecommerce platform, analytics and payment tools a site uses.
Logo priority: schema.org `logo` first, then the largest apple-touch-icon; `ogImage` and `favicon` are returned separately.

### Detect the email provider of a company domain

The actor reads the domain's MX records and maps them to a provider: Google Workspace, Microsoft 365, Zoho Mail,
Proton Mail, Mimecast, Proofpoint, Broadcom Email Security, GoDaddy, Fastmail, iCloud Mail, Yandex Mail, OVH, ImprovMX,
Amazon SES, Cisco Secure Email, Cloudflare Email Security / Email Routing, Barracuda, Rackspace Email, Tuta, Namecheap Private
Email, Titan, Sophos, Forcepoint, Hornetsecurity, Alibaba Mail or Tencent Exmail. The highest-priority MX that matches wins
(a security gateway in front of Google Workspace is reported as the gateway). Any other mail host is returned as `Other`, and
domains without MX records get `null`.

### Use cases

- **CRM enrichment and cleanup:** add logos, social links and founding dates to accounts that only have a domain.
- **Lead scoring and segmentation:** filter by technology (e.g. Shopify or WooCommerce) and email provider.
- **Market maps and competitor lists:** build a sheet of companies with tech stack and domain age.
- **Product UIs:** show a company logo and description from just a domain.
- **AI agents:** give an agent a quick, structured company snapshot through the Apify MCP server.

### Input example

```json
{
  "domains": ["stripe.com", "https://www.notion.so", "hubspot.com"],
  "includeTechStack": true,
  "includeDomainInfo": true,
  "maxConcurrency": 10
}
```

- `domains` — company domains or URLs; duplicates are removed.
- `includeTechStack` — detect 130+ technologies (default true).
- `includeDomainInfo` — domain age via RDAP and email provider via MX (default true).
- `maxConcurrency` — sites processed in parallel, 1 to 25 (default 10).

### Output example

```json
{
  "domain": "hubspot.com",
  "name": "HubSpot",
  "description": "HubSpot’s customer platform includes all the marketing, sales, customer service ...",
  "logo": "https://www.hubspot.com/hubfs/HubSpot_Logos/HSLogo_color.svg",
  "favicon": "https://www.hubspot.com/hubfs/HubSpot_Logos/HubSpot-Inversed-Favicon.png",
  "linkedin": "https://www.linkedin.com/company/hubspot",
  "twitter": "https://x.com/HubSpot",
  "youtube": "https://youtube.com/user/HubSpot",
  "foundingDate": "2006",
  "city": "Cambridge", "region": "MA", "country": "US",
  "technologyNames": ["Hubspot CMS", "Cloudflare", "HSTS", "Google Tag Manager"],
  "emailProvider": "Google Workspace",
  "domainCreatedAt": "2005-02-06T20:02:28Z",
  "domainAgeYears": 21.6
}
```

Fields are `null` when the website doesn't publish them. Sites behind aggressive bot protection get a `warning`
(domain, tech and email-provider data are still returned where possible). The example is shortened.

### Pricing

**$0.002 per company** ($2 per 1,000) — profile, tech stack and domain data all included. No subscription.
Examples: 1,000 companies = $2.00, 5,000 companies = $10.00, a 200-account CRM list = $0.40.
Invalid domains and sites that can't be reached get an `error` row and are **not charged**.

### Limitations

- One homepage request per company (plus DNS/RDAP lookups) — fast and polite. Sites that render everything with
  JavaScript may return fewer social links.
- Address, phone, legal name and founding date only appear when the site publishes schema.org Organization markup.
- Only the homepage is read: no employee counts, funding, revenue or people data.
- Social links are those linked from the homepage or listed in its structured data; a company with no such links returns `null`.
- Tech detection is signature-based and can miss tools that leave no trace in the homepage or headers.
- Built and maintained with AI assistance. Problems or requests: use the Issues tab.

### FAQ

**How do I get a company's logo from its domain?**
Add the domain and read the `logo` field (schema.org logo, or the largest apple-touch-icon as fallback). `ogImage` and `favicon` are separate fields.

**Can I find a company's LinkedIn page from its website?**
Yes, if the homepage links to it or lists it in its structured data. The `linkedin` field returns the company, school or showcase page URL, otherwise `null`.

**How can I tell which email provider a company uses?**
Keep `includeDomainInfo` on. The `emailProvider` field maps the domain's MX records to providers such as Google Workspace or Microsoft 365.

**Does it return personal emails or employee names?**
No. It only returns company-level data published on the company's own website.

**How do I find all Shopify (or WordPress) sites in a list?**
Run the list with `includeTechStack` on, then filter `technologyNames` for the technology you care about.

**Why are some fields empty or a warning shown?**
The site may not publish that data, may render content with JavaScript, or may block automated requests (the row then carries a `warning` explaining why: the HTTP status, a bot-challenge page, or the page it redirected to). Challenge pages are never reported as the company name.

### Related tools

Other actors by the same developer (same flat pay-per-result pricing, no subscription):

- [App Store Reviews Scraper – Apple iOS, Multi-Country](https://apify.com/forevertools/apple-app-store-reviews)
- [Article Scraper & Text Extractor – Clean Markdown for LLM/RAG](https://apify.com/forevertools/article-extractor)
- [Job Postings & Career Page Scraper – Workday, Greenhouse, Lever](https://apify.com/forevertools/ats-company-jobs)
- [Bluesky Posts Scraper – Profiles & Threads, No Login](https://apify.com/forevertools/bluesky-posts)
- [DNS Lookup & WHOIS Domain Checker — SPF/DMARC, SSL Expiry](https://apify.com/forevertools/domain-whois-dns-ssl)
- [Email Validator & MX Record Checker – Bulk, Disposable, Role](https://apify.com/forevertools/email-syntax-mx-checker)
- [Bulk PageSpeed Insights, Lighthouse & Core Web Vitals](https://apify.com/forevertools/pagespeed-core-web-vitals)
- [PDF Extractor & Parser – Bulk PDF to Text with Metadata](https://apify.com/forevertools/pdf-to-text-extractor)
- [SEO Audit Crawler & Broken Link Checker – Website Health](https://apify.com/forevertools/website-seo-audit)
- [Sitemap Extractor & Bulk URL Status Checker](https://apify.com/forevertools/sitemap-url-status-checker)
- [Steam Reviews Scraper – Game Reviews & App Details](https://apify.com/forevertools/steam-reviews)
- [Website Technology Detector – Tech Stack, CMS & Analytics](https://apify.com/forevertools/website-tech-stack-detector)
- [Website Screenshot API – Bulk Full Page PNG, JPEG & PDF](https://apify.com/forevertools/website-screenshot)

### Integrations

Run it from the Apify API, a schedule, or no-code tools: the Apify apps for **Zapier**, **Make** and **n8n** can start any public actor ("Run Actor") and read its dataset. AI agents can call it through the **Apify MCP server**.

### Ready-made examples

One-click tasks you can run or copy and adapt:

- [Enrich SaaS company domains: social links, tech stack, age](https://apify.com/forevertools/company-website-enrichment/examples/saas-company-enrichment)
- [Enrich e-commerce brand websites: socials, platform, age](https://apify.com/forevertools/company-website-enrichment/examples/ecommerce-brand-enrichment)
- [Find LinkedIn, X and GitHub links for a list of companies](https://apify.com/forevertools/company-website-enrichment/examples/company-social-links-finder)

# Actor input Schema

## `domains` (type: `array`):

Company websites to enrich: bare domains (stripe.com) or URLs (https://www.notion.so/pricing); only the host is used. Duplicates are removed. One row per company: name, description, logo, social profile links, company phone/address when published as schema.org data, tech stack, email provider, domain creation date. Invalid or unreachable sites return a row with `error` and are not charged.

## `includeTechStack` (type: `boolean`):

130+ technologies: CMS, ecommerce, frameworks, analytics, chat, payments.

## `includeDomainInfo` (type: `boolean`):

Registration date via RDAP and email provider via MX records.

## `maxConcurrency` (type: `integer`):

Sites processed in parallel.

## Actor input object example

```json
{
  "domains": [
    "stripe.com",
    "notion.so",
    "hubspot.com"
  ],
  "includeTechStack": true,
  "includeDomainInfo": true,
  "maxConcurrency": 10
}
```

# Actor output Schema

## `results` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "domains": [
        "stripe.com",
        "notion.so",
        "hubspot.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("forevertools/company-website-enrichment").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "domains": [
        "stripe.com",
        "notion.so",
        "hubspot.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("forevertools/company-website-enrichment").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "domains": [
    "stripe.com",
    "notion.so",
    "hubspot.com"
  ]
}' |
apify call forevertools/company-website-enrichment --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,forevertools/company-website-enrichment"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/tsKdUFnSCEMBf2c8c/builds/Jvp925H5PtilZs1el/openapi.json
