SimilarWeb Email Scraper avatar

SimilarWeb Email Scraper

Pricing

from $1.49 / 1,000 results

Go to Apify Store
SimilarWeb Email Scraper

SimilarWeb Email Scraper

SimilarWeb Email Scraper extracts publicly indexed email addresses using targeted keywords, location filters, custom email domains, and exclusion terms. Build structured SimilarWeb contact datasets for business research, lead discovery, company research, and contact analysis.

Pricing

from $1.49 / 1,000 results

Rating

0.0

(0)

Developer

Email Scraper

Email Scraper

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

a day ago

Last modified

Categories

Share

SimilarWeb Email Scraper

SimilarWeb Email Scraper is an Apify Actor for discovering publicly indexed email addresses associated with SimilarWeb pages. It uses targeted keywords, optional geographic terms, configurable email-domain suffixes, and exclusion filters to identify relevant SimilarWeb search results and return structured contact records.

You provide one or more search keywords, optionally specify a country, state, or city, choose the email domains you want to target, and set an email collection limit for each keyword + domain combination. The Actor returns structured records containing the keyword, result title, description, URL, and discovered email address.

This makes the SimilarWeb Email Scraper useful for business research, contact discovery, market research, company research, lead research, and structured dataset creation based on publicly indexed information.

What Is a SimilarWeb Email Scraper?

A SimilarWeb Email Scraper is designed to find email addresses appearing in publicly indexed SimilarWeb search-result descriptions.

Instead of manually searching SimilarWeb-related pages for contact information, you can provide targeted search terms and let the Actor process multiple keyword and email-domain combinations.

The Actor supports optional location targeting, custom email-domain suffixes, and exclusion words or phrases. Results are delivered as structured dataset records that can be reviewed and used for research or other data workflows.

The quality and quantity of results depend on what is publicly indexed and available for the selected searches. The Actor does not guarantee that every requested email target will be available.

Key Features

FeatureDescriptionUser Benefit
Keyword-based searchSearch using one or multiple keywords or queriesTarget specific business or professional topics
Location filteringOptionally add a country, state, or cityNarrow searches geographically
Custom email domainsSpecify domains such as @gmail.com or @yahoo.comFocus collection on relevant email types
Per-combination limitSet the maximum number of emails for each keyword + domain combinationControl the depth of each search
Exclude wordsSkip descriptions containing selected words or phrasesReduce unwanted results
Duplicate preventionPreviously collected email addresses are not added againKeep the resulting dataset cleaner
Structured datasetResults contain consistent fieldsMake collected data easier to review and analyze
Multiple keywordsProcess several search terms in one runBuild broader search coverage

What Data Can You Extract?

The SimilarWeb Email Scraper returns five user-facing data fields for each collected email.

  • Keyword — The search keyword associated with the result.
  • Title — The title of the matching search result.
  • Description — The publicly indexed description or snippet associated with the result.
  • URL — The URL associated with the search result.
  • Email — The email address identified in the result description that matches one of the configured domain suffixes.

These fields provide both the discovered contact information and the surrounding search-result context, making it easier to understand where each email was found.

The dataset is therefore more than a simple list of email addresses. It also preserves the keyword, title, description, and URL associated with each result.

Why Use This Actor?

Manual contact research can require repeatedly testing different search terms, checking result descriptions, identifying matching email domains, and recording the information in a structured format.

The SimilarWeb Email Scraper automates this repetitive collection workflow.

You can define several related keywords instead of relying on one broad search term. For example, a business research workflow might use terms such as marketing manager, digital marketing, online business, or business consultant.

You can also combine those keywords with different email-domain suffixes and optional geographic terms. This gives you a configurable way to organize searches around the information you actually need.

Benefits

  • Automated email discovery — Reduce repetitive manual searching for publicly indexed email information.
  • Structured results — Receive consistent keyword, title, description, URL, and email fields.
  • Flexible targeting — Combine multiple search keywords with selected email domains.
  • Geographic refinement — Add a country, state, or city when location-specific research is needed.
  • Result filtering — Exclude descriptions containing unwanted words or phrases.
  • Research-friendly output — Keep contextual information alongside each discovered email.
  • Scalable search configuration — Configure several keyword and domain combinations within one run.
  • Duplicate control — The Actor avoids adding the same email address repeatedly during a run.

How to Use the SimilarWeb Email Scraper

The basic workflow is straightforward:

  1. Enter one or more keywords in the keywords field.
  2. Optionally enter a country, state, or city in location.
  3. Add the email-domain suffixes you want to search for in customDomains.
  4. Set maxEmails according to the desired collection depth.
  5. Optionally add unwanted words or phrases to excludeWords.
  6. Start the Actor.
  7. Review the resulting dataset containing the matching search context and email addresses.

For broader research, use several specific keywords rather than one very broad term. More targeted queries can help distinguish different types of businesses, professionals, or topics.

Input

The Actor requires the keywords field. All other configuration fields are optional.

FieldTypeRequiredDefaultDescription
keywordsArray of stringsYes["marketing manager", "online business"]Search keywords or queries used to identify relevant SimilarWeb results.
locationStringNo""Optional country, state, or city used to narrow the search.
customDomainsArray of stringsNo["@gmail.com"]Email-domain suffixes to target, such as @gmail.com, @yahoo.com, or @outlook.com.
maxEmailsIntegerNo5Maximum target for each keyword + domain combination. Allowed range is 1–10,000.
excludeWordsArray of stringsNo[]Words or phrases that cause matching result descriptions to be skipped.

The keywords field uses a list of strings, allowing multiple search terms in the same run.

The location field can be left empty when geographic filtering is not required.

For customDomains, enter the complete email suffix including the @ symbol. You can specify multiple domains.

The maxEmails value applies independently to each keyword + domain combination. For example, two keywords and three domains create six combinations, with the configured target applied to each combination.

For free users, the Actor limits the effective maxEmails configuration to 100 when a higher value or no value is supplied. The Actor's code applies this ceiling to the configured per-combination limit.

Input Example

{
"keywords": [
"marketing manager",
"digital marketing",
"online business"
],
"location": "United States",
"customDomains": [
"@gmail.com",
"@yahoo.com",
"@outlook.com"
],
"maxEmails": 20,
"excludeWords": [
"crypto",
"onlyfans"
]
}

Output

Each collected result is added to the Apify dataset as a structured record.

FieldDescription
keywordThe keyword used for the search that produced the result.
titleTitle of the matching search result.
descriptionDescription or search-result snippet associated with the result.
urlURL associated with the matching result.
emailEmail address extracted from the result description that matches the configured domain pattern.

The Actor can produce multiple records from different keyword and domain combinations. Duplicate email addresses are not added again when they have already been collected during the run.

The output provides context around each email, which can be useful when reviewing or validating collected contact records.

Output Example

{
"keyword": "marketing manager",
"title": "Example Marketing Company",
"description": "Marketing services and business information. Contact: example@gmail.com",
"url": "https://www.similarweb.com/example",
"email": "example@gmail.com"
}

The example above demonstrates the output structure. Actual titles, descriptions, URLs, and email addresses depend on the publicly indexed search results available for the configured inputs.

Use Cases

The SimilarWeb Email Scraper can support several research and data-collection workflows.

  • Business research — Find publicly indexed contact information associated with relevant SimilarWeb results.
  • Lead research — Build structured contact datasets around selected professional or business keywords.
  • Market research — Investigate businesses or topics using multiple targeted search terms.
  • Company research — Organize search results and contact information around specific business categories.
  • Industry research — Search across professional terminology, industries, or business roles.
  • Geographic research — Combine keywords with locations to focus searches on particular markets.
  • Dataset creation — Build structured datasets containing email addresses and their associated search context.
  • Data analysis — Use keyword, description, URL, and email fields for subsequent research and analysis.
  • Contact discovery — Identify publicly indexed email addresses matching selected domain suffixes.

Competitive Advantages

The Actor provides several practical configuration options in a single workflow.

  • Multiple keywords can be processed in one run.
  • Multiple email-domain suffixes can be configured.
  • Geographic filtering is available when location matters.
  • Exclusion words and phrases can filter unwanted descriptions.
  • The resulting dataset preserves search context instead of returning only email addresses.
  • Each keyword + domain combination has its own configured collection target.
  • Duplicate email addresses are prevented from being added repeatedly.

These capabilities make the Actor adaptable to different research scopes without requiring users to manually repeat the same search process.

Advantages

  • Clear input structure with one required field and several optional controls.
  • Supports targeted rather than single-query research.
  • Produces structured Apify dataset records.
  • Includes contextual search-result information with each email.
  • Supports custom email-domain targeting.
  • Provides configurable exclusion filtering.
  • Supports geographic search refinement.
  • Allows users to control collection depth through maxEmails.

Limitations

The Actor works with publicly indexed search-result information, so results depend on what is available for the selected keywords, domains, and optional location.

Important limitations include:

  • Not every SimilarWeb page or business will have a publicly indexed email address.
  • A requested maxEmails value does not guarantee that the target number will be found.
  • Narrow keywords or locations may produce fewer results.
  • The selected email-domain suffixes determine which email addresses can be collected.
  • Exclusion terms can intentionally remove otherwise matching results.
  • Search availability can change over time.
  • The Actor does not guarantee that every returned email address remains active or belongs to the intended organization.
  • Free users have an effective maxEmails ceiling of 100 when a higher or unspecified value is supplied.

Because the Actor depends on publicly indexed information, empty or partial results should be treated as a possible outcome rather than an indication that a specific number of contacts must exist.

Pros and Cons

ProsCons
Multiple keyword supportResults depend on publicly indexed information
Custom email-domain targetingNarrow searches may return few emails
Optional geographic filteringRequested limits do not guarantee matching results
Exclusion words and phrasesCollected email addresses should be independently validated when important
Structured contextual outputSearch coverage can vary over time
Duplicate email preventionOnly configured email-domain patterns are targeted

Comparison With Alternative Approaches

CapabilitySimilarWeb Email ScraperManual / Typical Alternative
Keyword-based researchSupportedUsually performed manually
Multiple search termsSupportedRequires repeated searches
Email-domain filteringConfigurableOften requires manual filtering
Location refinementSupportedDepends on the search workflow
Exclusion filteringSupportedUsually performed manually
Structured outputDataset fields for each resultMay require manual formatting
Duplicate preventionSupported during collectionOften requires separate cleanup
Search-result contextTitle, description, URL, and keyword includedMay need manual recording

This comparison describes workflow differences rather than claiming that one approach is universally better.

Best Practices

  • Start with a small maxEmails value to test the search configuration.
  • Use several specific keywords instead of relying only on a broad keyword.
  • Add related job titles, business categories, or professional terms when appropriate.
  • Use multiple email domains when you want broader email coverage.
  • Leave location empty when you need broad geographic coverage.
  • Use a broader location when a narrow geographic term produces limited results.
  • Add excludeWords carefully because a matching exclusion term causes the complete result description to be skipped.
  • Review the output before using collected contact information in important workflows.
  • Treat email addresses as publicly discovered data that may require independent validation.

Troubleshooting

Empty Results

Try broader or more specific keywords, remove an overly narrow location, or add additional email-domain suffixes. A search may also have limited publicly indexed information.

Fewer Emails Than maxEmails

The limit is a target rather than a guarantee. The selected searches may simply contain fewer matching publicly indexed email addresses.

Unexpectedly Missing Results

Check whether the result description contains one of the configured exclusion words or phrases. Also verify that the email uses one of the configured domain suffixes.

Location Produces Limited Coverage

Try a broader country, state, or city value, or leave location empty when geographic filtering is not required.

Invalid Input

Check that keywords, customDomains, and excludeWords are arrays of strings and that maxEmails is an integer between 1 and 10,000.

Partial Results

Partial collection can occur when suitable indexed results are limited or when a search does not provide enough matching information. Review the query configuration before increasing collection targets.

Frequently Asked Questions

What does the SimilarWeb Email Scraper do?

It searches for publicly indexed SimilarWeb results using your configured keywords and email-domain suffixes, then extracts matching email addresses from result descriptions.

What data does the SimilarWeb Email Scraper return?

Each record contains keyword, title, description, url, and email.

Can I use multiple keywords?

Yes. The keywords input is an array, so you can provide multiple search terms in a single run.

Can I target a specific country or city?

Yes. Enter a country, state, or city in the optional location field to narrow the search geographically.

Can I search for different email providers?

Yes. The customDomains field accepts multiple email-domain suffixes, such as @gmail.com, @yahoo.com, and @outlook.com.

How does maxEmails work?

The configured value is applied to each keyword + domain combination. It can be set from 1 through 10,000 in the Actor input schema. Free users have an effective ceiling of 100 when a higher or unspecified value is supplied.

Can I exclude unwanted results?

Yes. Add words or phrases to excludeWords. If a configured exclusion matches the result description, that description is skipped and no email is collected from it.

Why did the Actor return fewer emails than requested?

maxEmails controls the collection target; it does not guarantee that the requested number exists in publicly indexed results. Broader keywords, additional domains, or a broader location can increase potential coverage.

Is the output structured?

Yes. Results are stored as structured Apify dataset records with consistent fields for keyword, title, description, URL, and email.

Can duplicate email addresses appear?

The Actor keeps track of collected addresses and does not add the same email address repeatedly during the run.

NLP Keywords

  • SimilarWeb email scraper
  • SimilarWeb email extraction
  • SimilarWeb contact data
  • SimilarWeb email finder
  • SimilarWeb contact scraper
  • SimilarWeb business contacts
  • SimilarWeb lead data
  • SimilarWeb profile emails
  • SimilarWeb email collection
  • SimilarWeb data extraction
  • SimilarWeb contact discovery
  • SimilarWeb business research
  • SimilarWeb lead research
  • SimilarWeb email addresses
  • SimilarWeb structured data
  • SimilarWeb search results
  • SimilarWeb company research
  • SimilarWeb keyword search
  • SimilarWeb contact information
  • SimilarWeb dataset
  • SimilarWeb email scraper tool
  • scrape emails from SimilarWeb
  • SimilarWeb email extractor
  • SimilarWeb contact information scraper
  • SimilarWeb business email finder
  • SimilarWeb lead scraper
  • SimilarWeb public email scraper
  • SimilarWeb company email extraction
  • SimilarWeb email data collection
  • SimilarWeb contact discovery tool
  • SimilarWeb keyword email search
  • SimilarWeb business contact extractor
  • SimilarWeb profile email extraction
  • SimilarWeb email research
  • SimilarWeb lead generation research
  • SimilarWeb contact dataset
  • SimilarWeb email search automation
  • SimilarWeb company contact data
  • SimilarWeb business research scraper
  • SimilarWeb email address finder

Final Overview

The SimilarWeb Email Scraper provides a configurable way to discover publicly indexed email addresses associated with SimilarWeb search results. Users can supply multiple keywords, optionally narrow searches by location, select email-domain suffixes, configure collection targets, and filter descriptions using exclusion terms.

The resulting Apify dataset preserves the keyword, result title, description, URL, and matching email address, giving users useful context for business research, lead research, company research, market analysis, and structured contact-data workflows.

For better coverage, start with focused but varied keywords, use appropriate email domains, and adjust location filtering according to the research objective. Always review and validate important contact information before relying on it for downstream activities.

Contact me: Alphascraper69@gmail.com