# Contact Details Extractor - Email, Phone & Social from Website (`tinyrex/contact-details-extractor`) Actor

Extract company emails, phones and social links from user-supplied domains. Role vs personal email flags, same-host contact pages, robots.txt. ~$0.002/domain with contacts; empty/blocked/failed free.

- **URL**: https://apify.com/tinyrex/contact-details-extractor.md
- **Developed by:** [TinyRex](https://apify.com/tinyrex) (community)
- **Categories:** Lead generation, SEO tools, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.60 / 1,000 domain with contacts

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Contact Details Extractor: get company emails, phones & socials from any website

**Enrich a list of company domains with public contact details** — emails, phone numbers, LinkedIn/X/Facebook/Instagram/YouTube/GitHub/TikTok links and addresses — as clean JSON or CSV. Paste domains you already have (CRM export, lead list, competitor set); the Actor only reads the homepage and common contact/about pages on **that same host**. No search-engine crawl, no login, no paywall bypass. **~$0.002 per domain with a contact found; empty, blocked or failed domains are free.**

### What you get from the Contact Details Extractor

- **Emails with GDPR-aware flags** — every `mailto:` and on-page address, tagged `roleOrGeneric` (info@, sales@, hello@…) vs `personalLooking` (first.last-style), plus the page it came from.
- **Phones, social profiles and addresses** — `tel:` links, LinkedIn company pages, X/Twitter, Facebook, Instagram, YouTube, GitHub, TikTok, and obvious `<address>` / schema.org PostalAddress blocks.
- **Same-host only, robots.txt respected** — homepage + `/contact`, `/about`, `/impressum` and contact-looking links on the homepage; never follows off the registrable host. Plain HTTP, 512 MB, no browser.

### Sample output (JSON)

```json
{
  "domain": "apify.com",
  "finalUrl": "https://apify.com/",
  "title": "Apify: Full-stack web scraping and data extraction platform",
  "emails": [
    { "email": "hello@apify.com", "roleOrGeneric": true, "personalLooking": false, "sources": ["https://apify.com/"] }
  ],
  "phones": [],
  "socials": {
    "linkedin": ["https://www.linkedin.com/company/apifytech"],
    "twitter": ["https://x.com/apify"],
    "facebook": [],
    "instagram": [],
    "youtube": [],
    "github": ["https://github.com/apify"],
    "tiktok": []
  },
  "addresses": [],
  "emailsFound": 1,
  "phonesFound": 0,
  "socialsFound": 3,
  "hasContact": true
}
```

Download the dataset as **CSV, Excel or JSON**, or pipe it to Google Sheets, Make, Zapier, n8n, webhooks or the Apify API.

### How to configure the Contact Details Extractor

1. Paste **domains or URLs**, one per line (`example.com`, `https://company.com`).
2. Optional: raise *Max pages per domain* (default 8) or turn off *Respect robots.txt*.
3. Run. Domains with at least one contact field are charged; the rest are free.

#### Example input

```json
{
  "domains": ["example.com", "apify.com"],
  "maxPagesPerDomain": 8,
  "respectRobotsTxt": true
}
```

| Setting | Default | Notes |
|---|---|---|
| domains | (required) | Company sites you supply — not scraped from Google or directories |
| maxPagesPerDomain | 8 | Homepage + contact/about candidates on the same host |
| respectRobotsTxt | true | Skip disallowed paths |
| maxConcurrency | 3 | Domains in parallel (keep low for politeness) |

### Common use cases

- **Lead enrichment** — add public emails and LinkedIn URLs to a domain list from your CRM or from the [Tech Stack Detector](https://apify.com/tinyrex/tech-stack-detector).
- **Sales / agency research** — find role inboxes (hello@, sales@) before outreach.
- **Vendor due diligence** — confirm the social profiles a company lists on its own site.
- **Data cleaning** — refresh contact fields for websites you already work with.

### Pricing (failed = free)

Pay per event:

- **Domain with contact** (`domain-contact-found`): ~$0.002 when at least one email, phone, social or address is found.
- **Free:** no contact fields, unreachable hosts, blocked pages (4xx/5xx), robots-skipped paths.

Set a max cost per run; the Actor stops gracefully when it is reached. See the *Pricing* tab for tier discounts (Bronze/Silver/Gold).

### Privacy & GDPR

Extracted emails and phones **may include personal data**. You are the data controller for how you use the results. This Actor only fetches pages on domains **you** supply, flags role vs personal-looking emails as a helper (not a filter), and does **not** claim to strip all personal data. Prefer role/generic addresses for outreach and follow applicable law (including GDPR). Do not build unsolicited spam lists.

### Limitations

- Public HTML on the same host only. JS-rendered contact widgets may be missed (no headless browser).
- Datacenter IP blocks may need a proxy.
- Phone/address extraction is heuristic — spot-check before outreach.

### FAQ — Contact Details Extractor

**How do I extract emails from a website list?**
Paste the domains into `domains`, run the Actor, and download the dataset. Each row is one domain with `emails`, `phones`, `socials` and source URLs.

**Is scraping contact emails from websites legal?**
You only get addresses the company already publishes on its own site. Legality depends on how *you* use the data (GDPR, CAN-SPAM, local law). Prefer role inboxes and get consent where required. This tool does not bypass logins or scrape people-search databases.

**How is this different from Hunter, Apollo or Clearbit?**
Those products often enrich from third-party databases and guess patterns. This Actor reads **only the websites you supply**, returns the exact page each email was found on, and flags role vs personal-looking addresses. No account on those services needed.

**Will it find LinkedIn company pages?**
Yes — LinkedIn, X/Twitter, Facebook, Instagram, YouTube, GitHub and TikTok profile/channel URLs linked from the site (share/intent links are ignored).

**What if a domain has no contact info?**
You are not charged. The row still appears so you can see what was tried (`pagesScraped`, status codes).

**Does it work with AI agents / MCP?**
Yes. Connect Claude, Cursor, VS Code, n8n or any MCP client to `https://mcp.apify.com?tools=tinyrex/contact-details-extractor` and ask e.g. *"Get emails and LinkedIn URLs for apify.com and stripe.com"*.

### Related actors

- [Tech Stack Detector](https://apify.com/tinyrex/tech-stack-detector) — technologies + email/DNS provider for the same domain list
- [Sitemap Scraper](https://apify.com/tinyrex/sitemap-scraper) — every URL from XML sitemaps
- [Shopify Products Scraper](https://apify.com/tinyrex/shopify-products-scraper) / [WooCommerce Products Scraper](https://apify.com/tinyrex/woocommerce-products-scraper) — product export when the stack is ecommerce

Built by **TinyRex**. Questions or feature requests? Open an issue on the Actor page.

# Actor input Schema

## `domains` (type: `array`):

Websites to analyze. Bare domains (example.com) or full URLs. For each domain the Actor fetches the homepage plus common contact/about pages on the same host only. Duplicates are removed. One charge per domain where at least one contact field is found.

## `maxPagesPerDomain` (type: `integer`):

Homepage + common contact/about paths + contact-looking links found on the homepage. Cap keeps the crawl small and polite.

## `respectRobotsTxt` (type: `boolean`):

Skip paths disallowed by robots.txt for user-agent \* (default on).

## `maxConcurrency` (type: `integer`):

How many domains are processed in parallel. Keep it low to be polite.

## `pageConcurrency` (type: `integer`):

Parallel page fetches within one domain.

## `requestTimeoutSecs` (type: `integer`):

How long to wait for each page before giving up.

## `proxyConfiguration` (type: `object`):

Optional proxy, only needed for sites that block datacenter IPs.

## Actor input object example

```json
{
  "domains": [
    "example.com",
    "apify.com"
  ],
  "maxPagesPerDomain": 8,
  "respectRobotsTxt": true,
  "maxConcurrency": 3,
  "pageConcurrency": 2,
  "requestTimeoutSecs": 20,
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# Actor output Schema

## `contacts` (type: `string`):

Emails, phones, socials and addresses per domain.

## `summary` (type: `string`):

Totals for the run.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "domains": [
        "example.com",
        "apify.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("tinyrex/contact-details-extractor").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "domains": [
        "example.com",
        "apify.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("tinyrex/contact-details-extractor").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "domains": [
    "example.com",
    "apify.com"
  ]
}' |
apify call tinyrex/contact-details-extractor --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,tinyrex/contact-details-extractor"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/ZllgD2fadq6Fg5TKN/builds/aVt8JLOCW9yyPWl0S/openapi.json
