# Website Contact Details Scraper — Emails, Phones, Socials (`tacps126/contact-details`) Actor

Find e-mails, phone numbers, social profiles (LinkedIn, X, Facebook, Instagram, YouTube…) and addresses on any list of websites. Checks contact, imprint and about pages automatically. Pay only for websites where contacts are found.

- **URL**: https://apify.com/tacps126/contact-details.md
- **Developed by:** [Tapaswai Ashok Choudhary](https://apify.com/tacps126) (community)
- **Categories:** Lead generation, Marketing, Integrations
- **Stats:** 1 total users, 0 monthly users, 0.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.60 / 1,000 website with contacts

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Website Contact Details Scraper — Emails, Phones, Socials

Turn a list of websites into a contact list. For each website you get **e-mail addresses, phone numbers, social profiles** (LinkedIn, X/Twitter, Facebook, Instagram, YouTube, TikTok, GitHub), the **company name, description and postal address**, and the page where the contacts were found — as a clean, spreadsheet-ready table. Export to CSV or Excel, push to your CRM or Google Sheets, or call it as an API.

**Built for** sales and business-development teams building lead lists, marketers and agencies doing outreach, recruiters, and anyone enriching a list of company domains.

### Why this scraper

- **Finds contacts where they actually are.** It reads the page you give, then picks the pages most likely to list contacts — *Contact*, *Imprint / Impressum*, *Legal notice*, *About*, *Team* — instead of crawling blindly. Five pages per site finds most contacts.
- **Reads what simple scrapers miss.** E-mails from links and text, e-mails hidden by common anti-spam protection, and the company's own structured data (the name, e-mail, phone and address a site publishes for search engines).
- **Clean results, not noise.** Image file names, placeholder addresses (`name@example.com`), tracking ids and share buttons are filtered out. Phone numbers come in international format, duplicates removed, and the website's own e-mail addresses are listed first.
- **Pay only for results.** Websites where nothing is found are **free**. Dead or unreachable domains are skipped in a second instead of wasting run time.
- **Fast and cheap to run.** A small native program — no browser — that runs in 256 MB and checks many websites in parallel. Platform usage for a typical run is a fraction of a cent.
- **Handles rate limits for you.** When a site says “slow down” (HTTP 429) the Actor waits as long as asked and adapts its speed; timeouts and dropped connections are retried automatically.
- **Never loses or double-bills work.** Each website is saved as soon as it's done; if the platform moves your run to another server it resumes without charging twice, and it stops cleanly at your maximum charge.

### How to use it

1. Paste websites into **Websites** — domains (`apify.com`) or any page URL.
2. Optionally change **Pages per website** (default 5), tick **Only the website's own e-mail addresses**, or report only websites with an **e-mail** or a **phone number**.
3. Press **Start**. Download the results from the **Output** tab, or connect an integration.

#### Example input

```json
{
  "startUrls": ["apify.com", "zalando.de", "mailchimp.com"],
  "maxPagesPerSite": 5,
  "sameDomainEmailsOnly": false
}
```

### What you get

One row per website:

| Field | Description |
|-------|-------------|
| `domain` / `url` | The website and the page it was read from |
| `companyName` / `description` | Company name and the site's own description |
| `primaryEmail` / `emails` | Best e-mail (the site's own domain first) and every e-mail found |
| `primaryPhone` / `phones` | First phone number and every number found, in international format where the site gives it |
| `linkedin` / `twitter` / `facebook` / `instagram` / `youtube` / `tiktok` / `github` / `vimeo` | Social profile links |
| `address` | Postal address, when the site publishes one |
| `contactPageUrl` | The contact / imprint / about page where details were found |
| `pagesScraped` / `scrapedAt` | Pages checked and when |
| `source` | When reading another scraper's results: the original row's name, id, address, phone and link |

Example:

```json
{
  "domain": "apify.com",
  "url": "https://apify.com/",
  "companyName": "Apify",
  "primaryEmail": "support@apify.com",
  "emails": ["support@apify.com", "hello@apify.com"],
  "linkedin": "https://www.linkedin.com/company/apify",
  "twitter": "https://x.com/apify",
  "github": "https://github.com/apify",
  "tiktok": "https://www.tiktok.com/@apifytech",
  "address": "Na Příkopě 959/27, Prague, 11000, CZ",
  "contactPageUrl": "https://apify.com/contact",
  "pagesScraped": 5
}
```

### Add contacts to another scraper's results

Already scraping businesses — from **Google Maps**, a directory, or any list that includes websites? Let this Actor enrich those results automatically:

1. Open the other Actor (for example a Google Maps scraper) and go to its **Integrations** tab.
2. Choose **Add integration → Actor**, pick **Website Contact Details Scraper**, and trigger it when the run **succeeds**.
3. Done. After every run, each website in its results is checked for e-mails, phones and social profiles.

You can also run it by hand on any existing dataset: pick it under **…or websites from another scraper's results**.

The website column is found automatically (`website`, `url`, `domain` and similar; map, social-media and job-board links are skipped). If yours is called something else, set **Website column**. Every result carries a `source` object with the original row's name, id, address and phone, so you can join the contacts back to your list. Rows that share a website are checked once.

### Only new websites

Working through a growing lead list? Turn on **Only new websites** and re-run with the whole list — websites an earlier run already delivered are skipped (and not charged).

### Use it as an instant API

The Actor also runs in **Standby mode** — an always-ready endpoint that returns contacts straight in the response, ideal for enriching leads from your own app, CRM workflow or AI agent. Find the URL on the **Standby** tab:

```bash
curl "https://<standby-url>/?url=apify.com,zalando.de" \
  -H "Authorization: Bearer <YOUR_APIFY_TOKEN>"
```

Any input field works as a query parameter (`url` is shorthand; separate several websites with commas), or `POST` the full input as JSON. The response is `{"success": true, "count": 2, "items": [ … ]}`.

### Pricing

Pay per result: about **$2 per 1,000 websites with contacts found** (lower on higher Apify plans) plus a tiny start fee. Websites with nothing found, and unreachable ones, cost nothing. See the **Pricing** tab for exact rates, and set a maximum charge per run to cap spend.

On Apify's **free plan** each run returns up to 100 websites — plenty to try everything. Any paid Apify plan removes the limit.

### FAQ

**Why is a website missing from the results?** Nothing was found on it, or it couldn't be loaded — those are free and listed in the run's summary. Turn on **Include websites with no contacts** to get a row for every input.

**Some sites return fewer details than I expected.** A few websites block automated visitors or load contacts only with JavaScript. Try raising **Pages per website**, or turn on a proxy.

**Is it legal?** The Actor collects contact details that websites publish publicly. You're responsible for how you use them — follow the privacy and marketing laws that apply to you (e.g. GDPR, CAN-SPAM) before contacting anyone.

### Support

Found a problem or need a field that isn't there yet? Open an issue on the **Issues** tab with the website and what you expected — requests genuinely decide what gets built next.

# Actor input Schema

## `startUrls` (type: `array`):

Website addresses or domains, one per line — e.g. `apify.com`, `https://www.zalando.de`. Any page of the site works; the Actor finds the contact pages itself. Or leave empty and pick a dataset below.

## `datasetId` (type: `string`):

Pick a dataset (for example from a Google Maps or company-list scrape) and every row's website is checked. Filled in automatically when this Actor runs as an integration after another Actor. Each result keeps the row's name / id in `source` so you can match it back.

## `websiteField` (type: `string`):

Only if the websites aren't found automatically: the column that holds them (dots for nested fields, e.g. `company.website`). By default `website`, `url`, `domain` and similar columns are used; map, social and job-board links are skipped.

## `maxPagesPerSite` (type: `integer`):

How many pages to check on each website, starting with the given page and then the most promising ones (contact, imprint, about, team…). 5 finds most contacts; raise it for large sites.

## `maxItems` (type: `integer`):

Maximum number of websites to return.

## `sameDomainEmailsOnly` (type: `boolean`):

Keep only addresses on the website's domain (e.g. drop the web agency's or a Gmail address).

## `requireEmail` (type: `boolean`):

Report (and charge) a website only if at least one e-mail address was found.

## `requirePhone` (type: `boolean`):

Report (and charge) a website only if at least one phone number was found.

## `includeEmpty` (type: `boolean`):

Also list websites where nothing was found (free of charge), so every input gets a row.

## `onlyNew` (type: `boolean`):

Skip websites an earlier run with the same input has already delivered — handy when you re-run a growing list.

## `monitorName` (type: `string`):

Optional name for the “only new” history, so several lists can run side by side.

## `proxyConfiguration` (type: `object`):

Not needed for most websites. Turn on for large lists or websites that limit repeated visits.

## `tenantId` (type: `string`):

Optional label used to keep separate histories for different clients or teams.

## `debug` (type: `boolean`):

Save extra diagnostics with the run (for support requests).

## Actor input object example

```json
{
  "startUrls": [
    "apify.com",
    "zalando.de",
    "mailchimp.com"
  ],
  "maxPagesPerSite": 5,
  "maxItems": 1000,
  "sameDomainEmailsOnly": false,
  "requireEmail": false,
  "requirePhone": false,
  "includeEmpty": false,
  "onlyNew": false,
  "proxyConfiguration": {
    "useApifyProxy": false
  },
  "debug": false
}
```

# Actor output Schema

## `contacts` (type: `string`):

One row per website with its contact details.

## `summary` (type: `string`):

Counts, websites that couldn't be loaded, and whether the run stopped early.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        "apify.com",
        "zalando.de",
        "mailchimp.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("tacps126/contact-details").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "startUrls": [
        "apify.com",
        "zalando.de",
        "mailchimp.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("tacps126/contact-details").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    "apify.com",
    "zalando.de",
    "mailchimp.com"
  ]
}' |
apify call tacps126/contact-details --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,tacps126/contact-details"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/gmYGde37Ci1IBEKZR/builds/aMGjJ3xzjxDqg8Re6/openapi.json
