# Unsplash Email Scraper - Domain Filter & Deduplication (`insightflow/unsplash-email-scraper-powerful-and-intelligent`) Actor

🌄 Unsplash Email Scraper — turn Unsplash searches into a clean photographer email list. Filter by keyword, location and domain, merge aliases and decode hidden addresses. 📷 For stock licensing & creative talent sourcing.

- **URL**: https://apify.com/insightflow/unsplash-email-scraper-powerful-and-intelligent.md
- **Developed by:** [InsightFlow](https://apify.com/insightflow) (community)
- **Categories:**
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.99 / 1,000 results

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

### Unsplash Email Scraper 🔍

**Unsplash Email Scraper** helps you extract emails from Unsplash publicly available sources faster and more consistently than manual browsing. It’s built for marketers, recruiters, sales pros, and researchers who need Unsplash lead generation, Unsplash contact scraper workflows, and a practical Unsplash email finder for outreach lists. ⚡

### 🌟 Key Features of Unsplash Email Scraper

| Feature | Benefit |
|---|---|
| ✅ **Targeted Keyword Search** | Uses your chosen keywords to focus on the most relevant Unsplash profiles and contributors. |
| ✅ **Custom Email Domain Filters** | Lets you narrow results to domains like `@gmail.com` or `@yahoo.com` for better lead quality. |
| ✅ **Location Filter** | Adds an optional location filter to help refine results for region-specific outreach. |
| ✅ **Max Email Cap** | Stops once your email limit is reached, helping you control run time and costs. |
| ✅ **Real-Time Dataset Saving** | Saves each result as it is found, so you don’t lose progress on longer runs. |
| ✅ **Resume-Friendly Progress Tracking** | Keeps track of progress so interrupted runs can continue more smoothly. |
| ✅ **Built-In Proxy Support** | Includes proxy support for more reliable scraping from publicly available web data. |

### 📥 Input — Unsplash Email Scraper Parameters

```json
{
  "keywords": [
    "manager",
    "founder"
  ],
  "location": "",
  "customDomains": [
    "@gmail.com",
    "@yahoo.com"
  ],
  "maxEmails": 20
}
```

| Parameter | Type | Required | Default | Description |
|---|---|---|---|---|
| `keywords` | Array | ✅ Yes | `["manager","founder"]` | A list of keywords or queries to find relevant Unsplash profiles and contact opportunities. |
| `location` | String | No | `""` | Optional location filter to narrow results by place. |
| `customDomains` | Array | No | `["@gmail.com","@yahoo.com"]` | A list of email domains to match, such as personal or company domains. |
| `maxEmails` | Integer | No | `20` | The maximum number of emails to collect before the actor stops. |

### 📤 Output — What Unsplash Email Scraper Returns

The actor saves each result as a JSON record in your Apify dataset.

```json
[
  {
    "keyword": "manager",
    "title": "Sarah Johnson",
    "description": "Portrait and documentary photographer based in London, United Kingdom.",
    "url": "https://unsplash.com/@sarahjohnson",
    "email": "sarah.johnson@gmail.com"
  }
]
```

| Field | Label | Format | Description |
|---|---|---|---|
| `keyword` | Keyword | text | The keyword that led to this result. |
| `title` | Title | text | The profile title returned in the dataset row. |
| `description` | Description | text | The public description or summary associated with the result. |
| `url` | Url | link | The public profile URL for the result. |
| `email` | Email | text | The email address extracted from publicly available data. |

### 💻 How to Use Unsplash Email Scraper — Step-by-Step

1. **Open the Actor** — Find **Unsplash Email Scraper** in Apify and open the actor page.
2. **Enter Keywords** — Add one or more keywords to target the profiles you want.
3. **Set Location** — Optionally narrow results by location for more focused Unsplash data extraction.
4. **Add Email Domains** — Include one or more domains to improve Unsplash contact extraction quality.
5. **Set Max Emails** — Choose how many emails you want the actor to collect.
6. **Run the Actor** — Start the run and monitor progress in the logs.
7. **Export Results** — Download your Unsplash outreach list from the dataset when the run finishes.

No coding required — just set your filters and let the actor work. 🚀

### 💡 Best Use Cases for Unsplash Email Scraper

- 🎯 **Unsplash Lead Generation** — Build a targeted outreach list from public Unsplash profiles and contributors.
- 📣 **Email Marketing** — Collect contact details for newsletter, partnership, or cold outreach workflows.
- 🔬 **Unsplash Profile Scraper** — Research public profiles and identify relevant creators for campaigns.
- 🖼️ **Unsplash Image Metadata Extraction** — Support content and creator analysis with public profile context.
- 🤝 **Unsplash Photographer Email Finder** — Find email contacts for photographers and visual creators on Unsplash.

### Disclaimer

This actor only accesses **publicly available data** on Unsplash. It does not scrape private profiles, authenticated content, or password-protected pages. Users are solely responsible for ensuring their use complies with Unsplash’s Terms of Service, GDPR, CCPA, and applicable anti-spam laws. This tool is intended for legitimate purposes only — lead generation, research, and marketing in compliance with local regulations. For data-removal requests, contact 📧 <insightflowofficial@gmail.com>.

### 🆘 Support & Feedback

Have a question or found an issue with the **Unsplash Email Scraper**? We’re here to help.

- 🐞 **Bug Reports:** Share the issue with our team so we can investigate it.
- ✨ **Custom Solutions & Feature Requests:** Reach out if you need tailored scraping workflows.
- 📧 **Email:** <insightflowofficial@gmail.com>

Your feedback shapes future improvements — we read every message.

### MX Lookup

Every address is checked at the DNS level: the actor resolves the mail domain's
MX records and reports what it found.

| Field | Meaning |
| --- | --- |
| `mxFound` | `true` when the domain publishes at least one mail server |
| `mxHost` | The lowest-preference (primary) mail server |
| `mxRecords` | Every MX record found, in preference order |
| `mxProvider` | Who runs the mail: Google Workspace, Microsoft 365, Zoho, Proton, ... |
| `mxStatus` | `found`, `no_records`, `no_such_domain`, `timeout`, `error`, or `skipped` |

**Inputs**

- **MX Lookup** - turn the check on or off (default: on).
- **Only keep contacts whose domain has an MX record** - drop unreachable domains.
  Only a definite negative (`no_records` / `no_such_domain`) drops a contact; a
  timeout or resolver error is treated as unknown and the lead is kept.
- **MX lookup timeout (seconds)** - per-domain DNS budget, 1-15s.

This is a domain-level check. It confirms the domain can receive mail; it does
not verify that an individual mailbox exists.

# Actor input Schema

## `keywords` (type: `array`):

A list of keywords or queries to search for.

## `location` (type: `string`):

Location to filter search results.

## `customDomains` (type: `array`):

List of custom email domains

## `maxEmails` (type: `integer`):

Maximum number of emails to collect. The scraper will stop once this limit is reached. Setting a higher limit allows for more potential results but doesn't guarantee reaching that number. This helps save costs by controlling scraping time.

## `minLeadScore` (type: `integer`):

Drop any contact whose Lead Score (0-100) falls below this. Leave at 0 for no filter.

## `mxLookup` (type: `boolean`):

Resolve each address's domain to its mail servers. Adds the resolved MX hosts and the mail provider (Google Workspace, Microsoft 365, ...) to every result, and feeds the Lead Score. DNS-level only - it confirms the domain can receive mail, not that the individual mailbox exists.

## `requireMxRecord` (type: `boolean`):

Drop any contact whose domain provably accepts no mail. A lookup that times out or errors is treated as unknown and kept, so a DNS hiccup never silently deletes good leads.

## `mxTimeoutSecs` (type: `integer`):

How long to wait for each DNS answer. Clamped to 1-15 seconds.

## Actor input object example

```json
{
  "keywords": [
    "photographer",
    "creative studio"
  ],
  "location": "",
  "customDomains": [
    "@gmail.com",
    "@yahoo.com"
  ],
  "maxEmails": 20,
  "minLeadScore": 0,
  "mxLookup": true,
  "requireMxRecord": false,
  "mxTimeoutSecs": 3
}
```

# Actor output Schema

## `dataset` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "keywords": [
        "photographer",
        "creative studio"
    ],
    "location": "",
    "customDomains": [
        "@gmail.com",
        "@yahoo.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("insightflow/unsplash-email-scraper-powerful-and-intelligent").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "keywords": [
        "photographer",
        "creative studio",
    ],
    "location": "",
    "customDomains": [
        "@gmail.com",
        "@yahoo.com",
    ],
}

# Run the Actor and wait for it to finish
run = client.actor("insightflow/unsplash-email-scraper-powerful-and-intelligent").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "keywords": [
    "photographer",
    "creative studio"
  ],
  "location": "",
  "customDomains": [
    "@gmail.com",
    "@yahoo.com"
  ]
}' |
apify call insightflow/unsplash-email-scraper-powerful-and-intelligent --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,insightflow/unsplash-email-scraper-powerful-and-intelligent"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/sBcp5gOExhVAupJp0/builds/KKeLE7hkqb8HdF8WX/openapi.json
