Tmall Email Scraper avatar

Tmall Email Scraper

Pricing

from $2.49 / 1,000 results

Go to Apify Store
Tmall Email Scraper

Tmall Email Scraper

Tmall Email Scraper SD - Tmall Email Scraper is a lead generation tool that extracts leads with public contact emails, account names and profile URLs from Tmall results by keyword, location and email domain - Tmall email extractor.

Pricing

from $2.49 / 1,000 results

Rating

0.0

(0)

Developer

Leads Scraper

Leads Scraper

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

5 days ago

Last modified

Categories

Share

Tmall Email Scraper — Read This Before You Run It

The Tmall Email Scraper searches Google's public index for Tmall pages that carry a contact email, and returns whatever it finds as clean, structured data.

Start with the honest part, because it decides whether this Actor is right for you: Tmall is largely excluded from Google's index, so expect very low volume.

In a measured live test the Tmall Email Scraper reached Google successfully, parsed correctly, and found exactly one result block to work with across the whole run.

That is not a bug in the Tmall Email Scraper, and it is not a proxy problem. It is what Google actually holds for tmall.com.

Everything below is written so you can decide with your eyes open. If you need volume today, the sibling Actors listed further down will serve you far better.

Important: the Tmall Email Scraper never logs into Tmall, never uses a Tmall API and never opens the site. Every record comes from publicly indexed Google search results.


Why the Tmall Email Scraper Returns So Little: Google vs Baidu

The reason is straightforward and worth stating plainly rather than dressing up.

Google is not the dominant search index for Chinese consumer marketplaces. Baidu is, and Baidu is where Tmall storefront and product pages are actually crawled, ranked and surfaced.

Google's coverage of tmall.com is thin by comparison, so a site:tmall.com query has very little corpus to draw on no matter how the query is phrased.

The Tmall Email Scraper is a Google-based tool by design. It queries through the Apify GOOGLE_SERP proxy, parses Google's result blocks, and cannot reach an index it does not read.

What that means in practice

Query expansion, extra keywords and higher page caps will not manufacture pages that Google never indexed. They widen the net over an ocean that is mostly empty here.

So treat the Tmall Email Scraper as a completeness tool: cheap to run, occasionally useful for a specific brand name, and never the backbone of a China sourcing list.

If your goal is a working Chinese supplier email list, the honest recommendation is to use the DHgate Email Scraper or the AliExpress Email Scraper instead — both platforms publish export-facing pages that Google indexes properly.


Key Features of the Tmall Email Scraper

The table below reflects what the Tmall Email Scraper genuinely does. The engine is solid; the corpus it searches is the constraint.

FeatureWhat it means in practice
Google site: data collectionQueries are restricted to tmall.com, so only Tmall pages are parsed
Query expansionEach keyword × domain pair is searched in several phrasings: base, quoted, intitle:, plus one variant per query modifier
Domain-filtered extractionOnly emails ending in your customDomains list are kept
Global deduplicationOne email appears once across every query and every page of the run
Subdomain handle recoveryTmall gives each shop a name.tmall.com host, so when Google exposes one the parser recovers it as username
Obfuscation-aware parserUnderstands name [at] domain [dot] com, name (at) domain, name @ domain.com, domain .com, zero-width characters and the full-width @
Junk filterRejects placeholders such as email@, yourname@, test@, xxx@ and single-character locals
Boundary-correct matching@gmail.com will not match inside @gmail.company or @gmail.com.br
Soft-wrap repairDiscards a hit that is only the tail of another email in the same result block
Structural HTML parsingLocates the <h3> title, then the smallest surrounding block — it does not depend on Google's CSS class names
Whole-page fallback parserIf Google's markup changes, the run degrades to "emails without shop metadata" rather than "no emails"
ConcurrencyAn asyncio worker pool runs several queries in parallel with a shared stop signal on maxEmails
Retry logic and proxy rotationUp to 3 attempts per page with exponential backoff and a fresh proxy session per request
Block detectionCAPTCHA, "unusual traffic" and consent pages are detected and retried, not silently counted as empty
Blocked-query requeueFailed or blocked queries are re-queued once at the end of the run
Resumable stateProgress is stored in the key-value store keyed by a hash of your input, saved on PERSIST_STATE, MIGRATING and ABORTING
Streaming dataset writesEach lead is pushed to the Apify dataset the moment it is found
Run summary loggingPages fetched, blocked pages, retries and emails per page are reported at the end

Block detection matters here more than on most siblings. It is what lets you tell "Google is blocking us" apart from "Google has nothing", and on Tmall the answer is almost always the second one.


How the Tmall Email Scraper Works: Crawler and Parser Pipeline

The pipeline is short and deliberately transparent. There is no browser, no JavaScript rendering, no authentication and no cookies.

1. Read input. The Tmall Email Scraper loads your keywords, optional location, email domains and limits.

2. Build queries. It composes ordinary Google queries with the site: operator, for example site:tmall.com supplier "@gmail.com".

3. Fetch search result pages. Pagination runs asynchronously with aiohttp through the Apify GOOGLE_SERP proxy, with retry logic and a fresh proxy session on every attempt.

4. Parse result blocks. For each result the Tmall Email Scraper finds the <h3> title, walks up to the smallest enclosing block, and reads the title, the breadcrumb host and the snippet.

5. Extract and normalise emails. A domain-filtered regex pulls addresses out of the block text, then normalisation and the junk filter clean them up.

6. Deduplicate and push. Every unique email is written to the dataset immediately.

Steps 1 to 5 work exactly as they do on high-yield siblings. Step 2 is simply pointed at a domain Google barely covers.


What Structured Data Does the Tmall Email Scraper Extract?

The Tmall Email Scraper produces one dataset item per unique email, and every item carries the same 14 fields.

Alongside the address you get whatever account label Google printed, the shop URL if one was exposed, and the snippet the email came from.

You also get full provenance: the keyword and the exact Google query that produced the lead, plus a UTC timestamp.

On Tmall that provenance is unusually valuable, because with so few results you want to know precisely which phrasing worked.

Nothing outside these 14 fields is collected. The Tmall Email Scraper does not touch orders, pricing, reviews or anything behind a login.


Tmall Email Scraper Input Schema

Every field below is taken verbatim from the Tmall Email Scraper input schema.

FieldTypeDefaultDescription
keywordsarray (required)["brand", "supplier"]Search terms describing the Tmall accounts you want (niche, job title, industry)
locationstring""Optional location phrase added to every query
customDomainsarray["@gmail.com", "@yahoo.com"]Only emails on these domains are collected; the leading @ is optional
maxEmailsinteger (1–10000)20Stop once this many unique emails have been collected
countryCodestring""Two-letter country code for the search proxy (US, GB, DE…)
expandQueriesbooleantrueSearch each keyword × domain pair with several phrasings
queryModifiersarray["email", "contact", "wholesale", "cooperation", "business"]Extra words combined with each keyword when expansion is on
maxPagesPerQueryinteger (1–50)30Page cap per query
maxConcurrencyinteger (1–20)5How many queries run in parallel

JSON input example

{
"keywords": ["brand", "supplier", "flagship store"],
"location": "",
"customDomains": ["@gmail.com", "@yahoo.com", "@outlook.com"],
"maxEmails": 50,
"countryCode": "US",
"expandQueries": true,
"queryModifiers": ["email", "contact", "wholesale", "cooperation", "business"],
"maxPagesPerQuery": 30,
"maxConcurrency": 5
}

Keep maxEmails modest. On this platform a high ceiling does not increase yield; it only lengthens a run that will finish empty-handed either way.

Why query expansion still matters

Google caps a single query at roughly 300 results, and expansion is the standard workaround across this Actor family.

On Tmall expansion is worth leaving on because every extra phrasing is another chance at the handful of indexed pages that exist. It cannot conjure new ones.

Because deduplication is global, expansion can only ever add rows — it never duplicates what you already have.


Tmall Email Scraper Output Schema

Every dataset item produced by the Tmall Email Scraper contains all 14 fields below.

FieldMeaning
networkPlatform name
keywordThe keyword that produced the lead
queryThe exact Google query used
titleRaw result title
accountNameAccount label Google prints (handle, display name, or shop label)
fullNameDisplay name parsed from a profile-style title; empty for page titles that carry no name
usernameURL-safe handle when the platform exposes one; otherwise null
profileUrlCanonical account URL when a handle is known; otherwise empty
urlDirect platform link when exposed, else the profile URL
descriptionBio or page snippet, cleaned of labels and engagement counters
emailLower-cased email address
emailDomainThe matched domain (e.g. @gmail.com)
possiblyTruncatedtrue when Google's snippet ellipsis touched the email — verify before sending
foundAtISO 8601 UTC timestamp

JSON output example — the realistic low-yield shape

The example below shows what a row looks like when Google surfaces a Tmall page but exposes no shop handle, which is the common case here.

{
"network": "Tmall",
"keyword": "supplier",
"query": "site:tmall.com supplier \"@gmail.com\" cooperation",
"title": "Tmall Global - brand cooperation",
"accountName": "Tmall",
"fullName": "",
"username": null,
"profileUrl": "",
"url": "",
"description": "Brand cooperation and wholesale enquiries: brandcoop.export@gmail.com",
"email": "brandcoop.export@gmail.com",
"emailDomain": "@gmail.com",
"possiblyTruncated": false,
"foundAt": "2026-08-31T09:42:17Z"
}

Note the empty profileUrl and the null username. When Google prints no name.tmall.com host, there is no handle to recover and those fields stay blank by design.

Export the Tmall Email Scraper dataset as JSON, CSV, XLSX or HTML, or pull it through the Apify API into your CRM.


How to Use the Tmall Email Scraper

Step 1 — open the Actor. Launch the Tmall Email Scraper on the Apify platform.

Step 2 — set a small maxEmails. Something like 20 to 50. There is no point paying for a long run against a thin index.

Step 3 — use specific brand or category keywords. A named brand you already know sells on Tmall gives you a better chance than a generic term.

Step 4 — widen your email domains. Add @outlook.com, @hotmail.com, @163.com and @qq.com alongside the defaults; Chinese sellers rarely use Gmail.

Step 5 — read the run summary. It reports pages fetched, blocked pages and emails per page, so you can see immediately whether Google returned anything at all.

If the summary shows pages fetched but no results, that is the expected outcome documented throughout this page — not a misconfiguration.


Use Cases for the Tmall Email Scraper

Use caseHow the Tmall Email Scraper helpsRealistic expectation
Named-brand contact lookupSearch a specific brand plus cooperation or businessBest-case scenario for this Actor
Cross-border sourcing researchConfirm whether a Tmall Global page carries a public addressOccasional hits
Coverage completenessAdd Tmall to a multi-platform sweep so the gap is documentedReliable, but usually returns little
Index auditingMeasure how much of tmall.com Google actually holds for your nicheWorks as intended
CRM enrichmentMatch any recovered shop label to accounts you already trackVery small volume
China market researchRead description snippets from whatever Google does surfaceThin sample

Notice that every row above is framed modestly. Anyone promising a large Tmall seller email list from Google search is not describing something that exists.


When to Use the Tmall Email Scraper — and When to Use a Sibling

Use the Tmall Email Scraper when you specifically need Tmall, accept the volume, and want the search done properly rather than by hand.

Use a sibling when you need contactable Chinese suppliers at scale. The strongest options in this family are:

If your real target is Western storefronts rather than Chinese marketplaces, the footprint-based Shopify Store Email Scraper and Wix Stores Email Scraper return merchant domains at a far higher rate.

The Taobao Email Scraper shares Tmall's indexing problem, since it sits on the same side of the Great Firewall.


Limitations of the Tmall Email Scraper

These are real constraints. The first one dominates everything else.

  • Very low volume by design of the index, not the Actor. Tmall is largely excluded from Google's index. A live test produced a single result block for the entire run. Baidu, not Google, is where these pages are properly indexed, and the Tmall Email Scraper does not read Baidu.
  • Only publicly indexed emails. If an address is not visible in Google's index, it cannot be found. Private data is never accessed.
  • Google's ~300-result cap. A single query returns roughly 300 results at most. Query expansion exists to work around this, though on Tmall the ceiling is rarely the binding constraint.
  • possiblyTruncated. When Google's snippet ellipsis touches an address, this flag is set to true. Verify those rows before sending anything.
  • Apify GOOGLE_SERP proxy required. The Actor cannot run without Apify proxy credentials.
  • Free-plan cap. Free Apify plans are limited to 100 emails per run. Paid plans are uncapped — a limit you are very unlikely to reach here.
  • username and profileUrl availability. These are only populated when Google's result exposes a name.tmall.com handle. Most Tmall rows show only a display label, leaving username as null and profileUrl empty. That is a Google limitation, not a bug.
  • Language and domain mismatch. Tmall sellers overwhelmingly use Chinese mailbox providers, so the default @gmail.com and @yahoo.com filters suppress most of what little exists.
  • No guaranteed volume. Results vary with keywords, domains and location, and on this platform the variance floor is zero.

Nothing here is affiliated with, endorsed by or officially supported by Tmall or Alibaba Group.


ActorWhat it collects
Tmall Email and Phone Number ScraperEmails and phone numbers from Tmall
Tmall Phone Number ScraperPublic phone numbers from Tmall
AliExpress Email ScraperPublic contact emails from AliExpress
Allegro Email ScraperPublic contact emails from Allegro
Amazon Email ScraperPublic contact emails from Amazon
Best Buy Seller Email ScraperPublic contact emails from Best Buy
BigCommerce Store Email ScraperPublic contact emails from BigCommerce Store
Cdiscount Email ScraperPublic contact emails from Cdiscount
Coupang Email ScraperPublic contact emails from Coupang
Depop Email ScraperPublic contact emails from Depop
DHgate Email ScraperPublic contact emails from DHgate
eBay Email ScraperPublic contact emails from eBay
Ecwid Store Email ScraperPublic contact emails from Ecwid Store
Etsy Email ScraperPublic contact emails from Etsy
Faire Email ScraperPublic contact emails from Faire
Flipkart Email ScraperPublic contact emails from Flipkart
Home Depot Seller Email ScraperPublic contact emails from Home Depot
JD.com Email ScraperPublic contact emails from JD.com
Lazada Email ScraperPublic contact emails from Lazada
Lowe's Seller Email ScraperPublic contact emails from Lowe's
MercadoLibre Email ScraperPublic contact emails from MercadoLibre
Mercari Email ScraperPublic contact emails from Mercari
Newegg Email ScraperPublic contact emails from Newegg
Otto Email ScraperPublic contact emails from Otto
Overstock Email ScraperPublic contact emails from Overstock
Poshmark Email ScraperPublic contact emails from Poshmark
Rakuten Email ScraperPublic contact emails from Rakuten
Shopee Email ScraperPublic contact emails from Shopee
Shopify Store Email ScraperPublic contact emails from Shopify Store
Taobao Email ScraperPublic contact emails from Taobao
Target Seller Email ScraperPublic contact emails from Target
Temu Email ScraperPublic contact emails from Temu
Vinted Email ScraperPublic contact emails from Vinted
Walmart Email ScraperPublic contact emails from Walmart
Wayfair Email ScraperPublic contact emails from Wayfair
Wish Email ScraperPublic contact emails from Wish
Wix Stores Email ScraperPublic contact emails from Wix Stores
WooCommerce Email ScraperPublic contact emails from WooCommerce
Zalando Email ScraperPublic contact emails from Zalando
AliExpress Email and Phone Number ScraperEmails and phone numbers from AliExpress
Allegro Email and Phone Number ScraperEmails and phone numbers from Allegro
Amazon Email and Phone Number ScraperEmails and phone numbers from Amazon

Tmall Email Scraper Example Run

A sourcing consultant is compiling a China supplier map and wants Tmall represented rather than silently skipped.

They run the Tmall Email Scraper with keywords: ["brand", "supplier"], customDomains: ["@gmail.com", "@outlook.com"] and maxEmails: 30, expecting little.

The run finishes quickly. The summary logs pages fetched and no blocked pages, confirming Google was reached and simply had almost nothing to return.

They record the gap in their notes, then re-run the same keywords through the DHgate Email Scraper, which is where the usable supplier contacts actually come from.


Tmall Email Scraper FAQ

How many results should I expect from the Tmall Email Scraper?

Very few, and often none. In a measured live test the run surfaced a single result block. Plan for a handful of rows at best, and do not build a campaign around this Actor.

Why is the yield so low — is something broken?

No. Tmall is largely excluded from Google's index. Baidu is the dominant search engine for Chinese marketplaces and is where these pages are properly crawled. The Tmall Email Scraper reads Google only.

Will more keywords or more pages fix it?

No. Expansion widens the search over the same thin corpus. It is worth leaving on, but it cannot create index coverage that does not exist.

Which sibling should I use instead?

For Chinese suppliers, the DHgate, AliExpress and Temu Actors. For Western storefronts, the footprint-based Shopify, Wix and Squarespace Actors return merchant domains at a far higher rate.

Does the Tmall Email Scraper log into Tmall?

No. It never logs in, never uses a Tmall API and never opens the site. All data comes from publicly indexed Google search results.

Why are username and profileUrl usually empty?

Because Google rarely exposes a name.tmall.com host in its result blocks. Without that host there is no handle to recover, so username stays null and profileUrl stays empty.

Should I change the email domains?

Yes. Tmall sellers overwhelmingly use Chinese providers, so adding domains such as @163.com and @qq.com gives the Tmall Email Scraper more to match against than Gmail alone.

Do I need an Apify proxy?

Yes. The Tmall Email Scraper requires the Apify GOOGLE_SERP proxy and cannot run without Apify proxy credentials.

How many emails can it collect per run?

Free Apify plans are capped at 100 emails per run and paid plans are uncapped — but on this platform the index, not the plan, is your ceiling.

What does possiblyTruncated: true mean?

Google's snippet ellipsis touched the email, so the address may be cut off. Verify those rows before sending.

What happens if a run is interrupted?

Progress is saved in the key-value store, keyed by a hash of your input, and flushed on PERSIST_STATE, MIGRATING and ABORTING. Restarting resumes rather than starting over.

Is this affiliated with Tmall?

No. The Tmall Email Scraper is an independent tool and is not affiliated with, endorsed by or officially supported by Tmall or Alibaba Group.


Responsible Use and GDPR

Publicly visible does not mean consent to bulk marketing. Treat every address the Tmall Email Scraper returns as personal data belonging to a real business contact.

Follow GDPR, CAN-SPAM, PIPL and any other regime that applies to you: identify yourself, state why you are contacting them, honour opt-outs immediately and delete records on request.

With volumes this low, individual, well-researched outreach is the only sensible approach anyway.


Leave a review

If the Tmall Email Scraper saved you time, please leave a star rating and a short review on the Actor page.

Reviews are how other buyers judge whether a tool works, and they tell us which features to build next.

If something did not work, email neurodata.apify@gmail.com instead - bugs get fixed faster than they get complained about.

Support

Questions, bug reports or a custom build request? Email neurodata.apify@gmail.com and include your run ID so the Tmall Email Scraper logs can be checked quickly.