# Website Contact Finder: Emails, Phones & Socials (`labrat011/website-contact-finder`) Actor

Find emails, phone numbers and social profiles on any list of websites. Reads contact, about, team and imprint pages in many languages, decodes Cloudflare-protected and \[at]/\[dot] emails, validates phones, reads schema.org data, optional MX email check. One row per website.

- **URL**: https://apify.com/labrat011/website-contact-finder.md
- **Developed by:** [mick\_](https://apify.com/labrat011) (community)
- **Categories:** Lead generation, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.55 / 1,000 website scanneds

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

<img src="https://apify-image-uploads-prod.s3.us-east-1.amazonaws.com/wCP1WauwRX2Gr3Gir-actor-Y46hXxe8QpY8RKS4G-QT5Uw5AeaH-website-contact-finder.png" alt="Website Contact Finder: Emails, Phones & Socials logo" width="120">

## Website Contact Finder: Emails, Phones & Socials

Give it a list of websites and get back, for each one, the **emails, phone numbers and social profiles** listed on the site, with the page each one was found on. Built for lead generation, sales prospecting and enriching company lists. No API keys.

| At a glance | |
|---|---|
| **You give it** | A list of website URLs |
| **You get** | One row per website: emails, phone numbers, social profiles, contact page, address |
| **Price** | $0.005 per run + $0.00085 per website, $0.0017 per verified email (Free plan, lower on paid plans) |
| **Speed** | 10 websites in about 26 seconds |
| **Needs** | Nothing: no login, no API key, no proxy |

### What you get

- **More emails.** On the same 10 test websites, this actor found **36 emails; a popular alternative found 10**, and every one of those 10 was in our 36.
- **Smart page choice.** The homepage, then contact, about, team, impressum and legal pages in many languages (kontakt, contacto, impressum, a-propos...), then pricing, press, careers, FAQ and catering pages, then the rest of the site.
- **Hidden emails decoded:** Cloudflare-protected addresses and `name [at] domain [dot] com` style obfuscation.
- **Validated phone numbers,** normalized to international and E.164 format, **deduplicated**. The same number written two ways counts once.
- **Social profiles:** LinkedIn, X/Twitter, Facebook, Instagram, YouTube, TikTok, GitHub, Pinterest, Threads. Share buttons are ignored.
- **schema.org data:** organization name, address, email and phone from JSON-LD when the site publishes it.
- **Source page for every email and phone,** and whether the email is on the website's own domain.
- **Optional MX check:** confirms the email's domain accepts mail.
- **Fair billing:** websites that do not answer get a diagnostic row and are **not charged**.
- **Fast:** 10 websites, 10 pages each, in 26 seconds with $0.0005 of platform usage.

### Use cases

- **Lead generation.** Turn a list of company domains into emails, phones and LinkedIn pages.
- **Enrich Google Maps or directory exports.** Add emails to a list of local businesses that only have a website.
- **Sales prospecting.** Department addresses (sales@, press@, partnerships@) straight from the company's own site.
- **Agency outreach.** Find the right inbox for hundreds of sites at once.
- **Data cleaning.** Check which websites still answer and which email domains still accept mail.

### Example inputs

#### A list of company websites

```json
{ "urls": ["katzsdelicatessen.com", "https://www.patagonia.com", "basecamp.com"], "maxPagesPerSite": 10 }
```

#### Only company-domain emails, verified

```json
{ "urls": ["..."], "sameDomainEmailsOnly": true, "verifyEmails": true }
```

#### European businesses

```json
{ "urls": ["https://www.example.de", "https://www.example.fr"], "defaultCountry": "DE", "maxPagesPerSite": 15 }
```

`defaultCountry` is used to read local phone numbers written without a country code.

### Input

| Field | What it does |
|---|---|
| `urls` | Websites, one per line. Domains or URLs. |
| `maxPagesPerSite` | Pages per website, 1 to 50. Default 10. The price per website does not change. |
| `defaultCountry` | Two-letter country for local phone numbers. Default `US`. |
| `sameDomainEmailsOnly` | Drop emails on other domains (gmail.com, agencies, partners). |
| `verifyEmails` | MX check for each email's domain. Charged per email. |
| `maxConcurrentSites` | Websites scanned in parallel. Default 8. |

### Output

One row per website. A real row from a test run on 2026-09-26:

```json
{
  "website": "https://www.katzsdelicatessen.com",
  "finalUrl": "https://katzsdelicatessen.com/",
  "domain": "katzsdelicatessen.com",
  "status": "partial",
  "error": null,
  "title": "Katz's Delicatessen - Since 1888  - NYC's oldest deli",
  "emails": [
    "cs@katzsdelicatessen.com",
    "events@katzsdelicatessen.com",
    "press@katzsdelicatessen.com",
    "sales@katzsdelicatessen.com"
  ],
  "phones": [
    "+1 800-446-8364",
    "+1 212-254-2246",
    "+1 212-254-2247"
  ],
  "emailDetails": [
    {
      "email": "cs@katzsdelicatessen.com",
      "sameDomain": true,
      "foundOn": "https://katzsdelicatessen.com/faqs",
      "mxValid": true
    },
    {
      "email": "events@katzsdelicatessen.com",
      "sameDomain": true,
      "foundOn": "https://katzsdelicatessen.com/faqs",
      "mxValid": true
    },
    {
      "email": "press@katzsdelicatessen.com",
      "sameDomain": true,
      "foundOn": "https://katzsdelicatessen.com/",
      "mxValid": true
    },
    {
      "email": "sales@katzsdelicatessen.com",
      "sameDomain": true,
      "foundOn": "https://katzsdelicatessen.com/privacypolicy",
      "mxValid": true
    }
  ],
  "phoneDetails": [
    {
      "e164": "+18004468364",
      "formatted": "+1 800-446-8364",
      "foundOn": "https://katzsdelicatessen.com/"
    },
    {
      "e164": "+12122542246",
      "formatted": "+1 212-254-2246",
      "foundOn": "https://katzsdelicatessen.com/"
    },
    {
      "e164": "+12122542247",
      "formatted": "+1 212-254-2247",
      "foundOn": "https://katzsdelicatessen.com/events"
    }
  ],
  "contactPageUrl": "https://katzsdelicatessen.com/contact",
  "address": null,
  "organizationName": null,
  "linkedin": null,
  "twitter": "https://twitter.com/KatzsDeli",
  "facebook": "https://www.facebook.com/katzsdeli",
  "instagram": "https://www.instagram.com/katzsdeli",
  "youtube": null,
  "tiktok": null,
  "github": null,
  "pinterest": null,
  "threads": null,
  "pagesScanned": 8,
  "pagesFailed": 2
}
```

`status` is `succeeded`, `partial` (some pages failed), or `failed` (the site did not answer, not charged).

### Pricing

Pay per event:

- **$0.005** per run start
- **$0.00085** per website scanned (flat, however many pages)
- **$0.0017** per email checked, only when `verifyEmails` is on

Lower on higher Apify plans. Websites that do not answer are not charged. Apify platform usage is billed separately to your account and is tiny: 10 websites used $0.0005.

| Websites | Actor cost |
|---|---|
| 10 | about $0.014 |
| 100 | about $0.09 |
| 1,000 | about $0.86 |

### Automate it with n8n

Each workflow uses n8n's official **Apify** node, operation **Run actor and get dataset**, actor `labrat011/website-contact-finder`.

#### 1. Enrich a Google Sheet of companies

```
Schedule Trigger (daily) or Google Sheets Trigger (row added)
  > Google Sheets: read rows where email is empty
  > Aggregate: collect the website column into one list
  > Apify: Run actor and get dataset   { "urls": {{ JSON.stringify($json.website) }}, "sameDomainEmailsOnly": true }
  > Google Sheets: update each row with emails[0], phones[0], linkedin
```

Switch **Input JSON** to Expression so the list of websites is filled in.

#### 2. Google Maps leads to CRM

```
Manual Trigger
  > Apify: Run a Google Maps scraper (for example labrat011/google-maps-scraper)
  > Filter: website is not empty
  > Apify: Run actor and get dataset   (website-contact-finder with those websites)
  > Merge: by website
  > HubSpot / Pipedrive: create companies and contacts
```

#### 3. Inbound form enrichment

```
Webhook (new signup with a company domain)
  > Apify: Run actor and get dataset   { "urls": ["{{ $json.domain }}"] }
  > Slack: "New signup from {{ $json.title }}: {{ $json.phones[0] }} {{ $json.linkedin }}"
```

#### 4. Outreach list with verified inboxes

```
Manual Trigger
  > Apify: Run actor and get dataset   (your list, verifyEmails true)
  > Split Out: emailDetails
  > Filter: mxValid is true and sameDomain is true
  > Google Sheets / Instantly / Lemlist: add as leads
```

#### 5. Monthly website health check

```
Schedule Trigger (monthly)
  > Apify: Run actor and get dataset   (all client websites, maxPagesPerSite 1)
  > Filter: status is failed
  > Email: "These sites did not answer: ..."
```

### For AI agents

- **Actor:** `labrat011/website-contact-finder`
- **Smallest input:** `{ "urls": ["https://apify.com"] }`
- **One row = one website, with every email, phone and social profile found on it.** Key fields: `website`, `domain`, `status`, `emails`, `phones`, `contactPageUrl`, `linkedin`, `twitter`.
- **Billing:** `apify-actor-start` once per run, `website-scanned` per website that answers, `email-verified` per email only when `verifyEmails` is on. Cap spend with the number of `urls`, `maxPagesPerSite` or a maximum cost per run.
- **Run it:** `POST https://api.apify.com/v2/acts/labrat011~website-contact-finder/run-sync-get-dataset-items` with the input as the JSON body, or call it from the Apify MCP server.
- **Done signal:** the run's status message reads `Scanned N websites (F did not answer), found E emails.`

### FAQ

**Does it guess emails?** No. Every email and phone was published on the website. The `foundOn` field shows the page.

**What does the MX check prove?** That the email's domain is set up to receive mail. It does not prove the specific mailbox exists.

**Some websites return nothing.** Some sites block automated visitors or load contacts with JavaScript. Try turning on a proxy for those. Sites that do not answer at all are not charged.

### Support

Open an issue on the actor's Issues tab with the run ID and it will be looked at.

# Changelog

This Actor's version history is a separate document: https://apify.com/labrat011/website-contact-finder/changelog.md

# Actor input Schema

## `urls` (type: `array`):

Domains or URLs, one per line: acme.com, https://www.acme.com. Each site is scanned from its homepage.

## `maxPagesPerSite` (type: `integer`):

The homepage plus the most contact-like pages (contact, about, team, impressum, legal...). The price per website is the same whatever you set here.

## `defaultCountry` (type: `string`):

Two-letter country used to read local phone numbers written without a country code. Numbers with a + prefix are read as written.

## `sameDomainEmailsOnly` (type: `boolean`):

Drop gmail.com, agency and partner addresses.

## `verifyEmails` (type: `boolean`):

A DNS check per email domain. Charged per email checked.

## `maxConcurrentSites` (type: `integer`):

How many websites are scanned in parallel.

## `proxyConfiguration` (type: `object`):

Not needed for most sites. Proxy traffic is billed to your account by Apify.

## Actor input object example

```json
{
  "urls": [
    "https://apify.com",
    "https://www.katzsdelicatessen.com"
  ],
  "maxPagesPerSite": 10,
  "defaultCountry": "US",
  "sameDomainEmailsOnly": false,
  "verifyEmails": false,
  "maxConcurrentSites": 8,
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# Actor output Schema

## `contacts` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "urls": [
        "https://apify.com",
        "https://www.katzsdelicatessen.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("labrat011/website-contact-finder").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "urls": [
        "https://apify.com",
        "https://www.katzsdelicatessen.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("labrat011/website-contact-finder").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "urls": [
    "https://apify.com",
    "https://www.katzsdelicatessen.com"
  ]
}' |
apify call labrat011/website-contact-finder --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,labrat011/website-contact-finder"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/Y46hXxe8QpY8RKS4G/builds/TnnbeFBvWMXW1eH7b/openapi.json
