# Website Contact Details Extractor (`weio/website-contact-details-extractor`) Actor

Extract business contact details from a list of websites: role emails (info@, sales@, support@), phone numbers, social media links (Facebook, Instagram, LinkedIn, X, YouTube, TikTok) and the contact page URL. Business contacts only, no personal data. Email and contact scraper for B2B leads.

- **URL**: https://apify.com/weio/website-contact-details-extractor.md
- **Developed by:** [Weio, Inc.](https://apify.com/weio) (community)
- **Categories:** Lead generation, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$3.00 / 1,000 site with contact details

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Website Contact Details Extractor (business emails, phones, social links)

Give it a list of websites and get one row per site with the business contact details the site itself publishes:
business name, role emails (info@, sales@, office@ ...), phone numbers, social profile links and the contact page.
Useful for building B2B outreach lists, enriching a CRM, or checking that your own sites show current contact details.

### What it does

For each site the actor fetches the homepage over plain HTTP(S) (no browser), then at most two more pages on the same
domain: the contact page and an about page, if the homepage links to them. Results from the pages are merged.

- **emails**: role/business addresses only (info@, sales@, contact@, office@, support@ ...), lowercased and deduplicated.
- **phones**: from `tel:` links and phone-number patterns in the page (North American format; numbers can include false positives such as placeholder digits).
- **socials**: first link found for facebook, instagram, linkedin, x (x.com / twitter.com), youtube, tiktok.
- **business\_name**: the site's `og:site_name`, else the page title.
- **contact\_page\_url**, **pages\_checked**, and **error** (null, or a short reason such as `http 403` or `invalid domain`).

Sites that render their contact details only with JavaScript, or that block automated requests, will return few or no details.

### Input

```json
{
  "websites": ["zingermans.com", "https://anchorbrewing.com"],
  "onlyWithEmail": false,
  "concurrency": 8
}
```

- `websites`: domains or URLs, one per line, up to 5,000 per run.
- `onlyWithEmail`: if true, sites without a business email are left out of the results and not charged.
- `concurrency`: websites processed in parallel (1-32, default 8).

### Output example

```json
{
  "input": "zingermans.com",
  "domain": "zingermans.com",
  "final_url": "https://zingermans.com/",
  "business_name": "Zingerman's: Online Shopping for Food and Gifts",
  "emails": [
    "service@zingermans.com"
  ],
  "phones": [
    "+17344362006",
    "+17344776986",
    "+17344776988",
    "+18662606169",
    "+18886368162"
  ],
  "socials": {
    "facebook": "https://www.facebook.com/Zingermans/",
    "instagram": "https://www.instagram.com/zingermansmailorder/"
  },
  "contact_page_url": "https://zingermans.com/CustService.aspx",
  "pages_checked": [
    "https://zingermans.com/",
    "https://zingermans.com/CustService.aspx",
    "https://zingermans.com/AboutUs.aspx"
  ],
  "error": null
}
```

### Pricing

$3.00 per 1,000 sites ($0.003 per site), platform usage included. You are charged only for rows with no error and
at least one email, phone or social link. Rows with an error or with nothing found are returned free, and with
`onlyWithEmail` sites without a business email are skipped and free too.

### Privacy and limits

- Business/role email addresses only. Addresses that look like a person's name (john.smith@...) are dropped on purpose and never returned.
- Only publicly reachable pages are read: the homepage and at most two same-domain pages. No logins, no forms, no page text is stored or returned.
- Requests to private or internal network addresses are refused.
- The actor does not read or follow robots.txt. You are responsible for using the results lawfully, including any email-marketing and data-protection rules that apply to you.

### Related actors from Weio

- [Website Tech Stack Detector](https://apify.com/weio/website-tech-stack-detector): CMS, store platform, analytics, pixels and hosting of a list of websites.
- [Local Business Website Audit](https://apify.com/weio/local-business-website-audit): mobile layout, HTTPS, title/description and contact checks for small-business sites.
- [Domain WHOIS, DNS and Email Security Checker](https://apify.com/weio/domain-whois-dns-email-security-checker): registration, DNS, MX, SPF and DMARC for a list of domains.

# Actor input Schema

## `websites` (type: `array`):

Domains or URLs, one per line (max 5,000 per run).

## `onlyWithEmail` (type: `boolean`):

Skip (and do not charge for) sites where no role/business email was found.

## `concurrency` (type: `integer`):

How many websites are processed at the same time (1-32).

## Actor input object example

```json
{
  "websites": [
    "zingermans.com",
    "anchorbrewing.com"
  ],
  "onlyWithEmail": false,
  "concurrency": 8
}
```

# Actor output Schema

## `contactRows` (type: `string`):

The default dataset with one row per processed website.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "websites": [
        "zingermans.com",
        "anchorbrewing.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("weio/website-contact-details-extractor").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "websites": [
        "zingermans.com",
        "anchorbrewing.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("weio/website-contact-details-extractor").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "websites": [
    "zingermans.com",
    "anchorbrewing.com"
  ]
}' |
apify call weio/website-contact-details-extractor --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,weio/website-contact-details-extractor"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/dJERrAAPEQQbOlJwt/builds/6zl4MYGk5Hm7fbx0c/openapi.json
