Tmall Email Scraper
Pricing
from $2.49 / 1,000 results
Tmall Email Scraper
Tmall Email Scraper SD - Tmall Email Scraper is a lead generation tool that extracts leads with public contact emails, account names and profile URLs from Tmall results by keyword, location and email domain - Tmall email extractor.
Pricing
from $2.49 / 1,000 results
Rating
0.0
(0)
Developer
Leads Scraper
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
5 days ago
Last modified
Categories
Share
Tmall Email Scraper — Read This Before You Run It
The Tmall Email Scraper searches Google's public index for Tmall pages that carry a contact email, and returns whatever it finds as clean, structured data.
Start with the honest part, because it decides whether this Actor is right for you: Tmall is largely excluded from Google's index, so expect very low volume.
In a measured live test the Tmall Email Scraper reached Google successfully, parsed correctly, and found exactly one result block to work with across the whole run.
That is not a bug in the Tmall Email Scraper, and it is not a proxy problem. It is what Google actually holds for tmall.com.
Everything below is written so you can decide with your eyes open. If you need volume today, the sibling Actors listed further down will serve you far better.
Important: the Tmall Email Scraper never logs into Tmall, never uses a Tmall API and never opens the site. Every record comes from publicly indexed Google search results.
Why the Tmall Email Scraper Returns So Little: Google vs Baidu
The reason is straightforward and worth stating plainly rather than dressing up.
Google is not the dominant search index for Chinese consumer marketplaces. Baidu is, and Baidu is where Tmall storefront and product pages are actually crawled, ranked and surfaced.
Google's coverage of tmall.com is thin by comparison, so a site:tmall.com query has very little corpus to draw on no matter how the query is phrased.
The Tmall Email Scraper is a Google-based tool by design. It queries through the Apify GOOGLE_SERP proxy, parses Google's result blocks, and cannot reach an index it does not read.
What that means in practice
Query expansion, extra keywords and higher page caps will not manufacture pages that Google never indexed. They widen the net over an ocean that is mostly empty here.
So treat the Tmall Email Scraper as a completeness tool: cheap to run, occasionally useful for a specific brand name, and never the backbone of a China sourcing list.
If your goal is a working Chinese supplier email list, the honest recommendation is to use the DHgate Email Scraper or the AliExpress Email Scraper instead — both platforms publish export-facing pages that Google indexes properly.
Key Features of the Tmall Email Scraper
The table below reflects what the Tmall Email Scraper genuinely does. The engine is solid; the corpus it searches is the constraint.
| Feature | What it means in practice |
|---|---|
Google site: data collection | Queries are restricted to tmall.com, so only Tmall pages are parsed |
| Query expansion | Each keyword × domain pair is searched in several phrasings: base, quoted, intitle:, plus one variant per query modifier |
| Domain-filtered extraction | Only emails ending in your customDomains list are kept |
| Global deduplication | One email appears once across every query and every page of the run |
| Subdomain handle recovery | Tmall gives each shop a name.tmall.com host, so when Google exposes one the parser recovers it as username |
| Obfuscation-aware parser | Understands name [at] domain [dot] com, name (at) domain, name @ domain.com, domain .com, zero-width characters and the full-width @ |
| Junk filter | Rejects placeholders such as email@, yourname@, test@, xxx@ and single-character locals |
| Boundary-correct matching | @gmail.com will not match inside @gmail.company or @gmail.com.br |
| Soft-wrap repair | Discards a hit that is only the tail of another email in the same result block |
| Structural HTML parsing | Locates the <h3> title, then the smallest surrounding block — it does not depend on Google's CSS class names |
| Whole-page fallback parser | If Google's markup changes, the run degrades to "emails without shop metadata" rather than "no emails" |
| Concurrency | An asyncio worker pool runs several queries in parallel with a shared stop signal on maxEmails |
| Retry logic and proxy rotation | Up to 3 attempts per page with exponential backoff and a fresh proxy session per request |
| Block detection | CAPTCHA, "unusual traffic" and consent pages are detected and retried, not silently counted as empty |
| Blocked-query requeue | Failed or blocked queries are re-queued once at the end of the run |
| Resumable state | Progress is stored in the key-value store keyed by a hash of your input, saved on PERSIST_STATE, MIGRATING and ABORTING |
| Streaming dataset writes | Each lead is pushed to the Apify dataset the moment it is found |
| Run summary logging | Pages fetched, blocked pages, retries and emails per page are reported at the end |
Block detection matters here more than on most siblings. It is what lets you tell "Google is blocking us" apart from "Google has nothing", and on Tmall the answer is almost always the second one.
How the Tmall Email Scraper Works: Crawler and Parser Pipeline
The pipeline is short and deliberately transparent. There is no browser, no JavaScript rendering, no authentication and no cookies.
1. Read input. The Tmall Email Scraper loads your keywords, optional location, email domains and limits.
2. Build queries. It composes ordinary Google queries with the site: operator, for example site:tmall.com supplier "@gmail.com".
3. Fetch search result pages. Pagination runs asynchronously with aiohttp through the Apify GOOGLE_SERP proxy, with retry logic and a fresh proxy session on every attempt.
4. Parse result blocks. For each result the Tmall Email Scraper finds the <h3> title, walks up to the smallest enclosing block, and reads the title, the breadcrumb host and the snippet.
5. Extract and normalise emails. A domain-filtered regex pulls addresses out of the block text, then normalisation and the junk filter clean them up.
6. Deduplicate and push. Every unique email is written to the dataset immediately.
Steps 1 to 5 work exactly as they do on high-yield siblings. Step 2 is simply pointed at a domain Google barely covers.
What Structured Data Does the Tmall Email Scraper Extract?
The Tmall Email Scraper produces one dataset item per unique email, and every item carries the same 14 fields.
Alongside the address you get whatever account label Google printed, the shop URL if one was exposed, and the snippet the email came from.
You also get full provenance: the keyword and the exact Google query that produced the lead, plus a UTC timestamp.
On Tmall that provenance is unusually valuable, because with so few results you want to know precisely which phrasing worked.
Nothing outside these 14 fields is collected. The Tmall Email Scraper does not touch orders, pricing, reviews or anything behind a login.
Tmall Email Scraper Input Schema
Every field below is taken verbatim from the Tmall Email Scraper input schema.
| Field | Type | Default | Description |
|---|---|---|---|
keywords | array (required) | ["brand", "supplier"] | Search terms describing the Tmall accounts you want (niche, job title, industry) |
location | string | "" | Optional location phrase added to every query |
customDomains | array | ["@gmail.com", "@yahoo.com"] | Only emails on these domains are collected; the leading @ is optional |
maxEmails | integer (1–10000) | 20 | Stop once this many unique emails have been collected |
countryCode | string | "" | Two-letter country code for the search proxy (US, GB, DE…) |
expandQueries | boolean | true | Search each keyword × domain pair with several phrasings |
queryModifiers | array | ["email", "contact", "wholesale", "cooperation", "business"] | Extra words combined with each keyword when expansion is on |
maxPagesPerQuery | integer (1–50) | 30 | Page cap per query |
maxConcurrency | integer (1–20) | 5 | How many queries run in parallel |
JSON input example
{"keywords": ["brand", "supplier", "flagship store"],"location": "","customDomains": ["@gmail.com", "@yahoo.com", "@outlook.com"],"maxEmails": 50,"countryCode": "US","expandQueries": true,"queryModifiers": ["email", "contact", "wholesale", "cooperation", "business"],"maxPagesPerQuery": 30,"maxConcurrency": 5}
Keep maxEmails modest. On this platform a high ceiling does not increase yield; it only lengthens a run that will finish empty-handed either way.
Why query expansion still matters
Google caps a single query at roughly 300 results, and expansion is the standard workaround across this Actor family.
On Tmall expansion is worth leaving on because every extra phrasing is another chance at the handful of indexed pages that exist. It cannot conjure new ones.
Because deduplication is global, expansion can only ever add rows — it never duplicates what you already have.
Tmall Email Scraper Output Schema
Every dataset item produced by the Tmall Email Scraper contains all 14 fields below.
| Field | Meaning |
|---|---|
network | Platform name |
keyword | The keyword that produced the lead |
query | The exact Google query used |
title | Raw result title |
accountName | Account label Google prints (handle, display name, or shop label) |
fullName | Display name parsed from a profile-style title; empty for page titles that carry no name |
username | URL-safe handle when the platform exposes one; otherwise null |
profileUrl | Canonical account URL when a handle is known; otherwise empty |
url | Direct platform link when exposed, else the profile URL |
description | Bio or page snippet, cleaned of labels and engagement counters |
email | Lower-cased email address |
emailDomain | The matched domain (e.g. @gmail.com) |
possiblyTruncated | true when Google's snippet ellipsis touched the email — verify before sending |
foundAt | ISO 8601 UTC timestamp |
JSON output example — the realistic low-yield shape
The example below shows what a row looks like when Google surfaces a Tmall page but exposes no shop handle, which is the common case here.
{"network": "Tmall","keyword": "supplier","query": "site:tmall.com supplier \"@gmail.com\" cooperation","title": "Tmall Global - brand cooperation","accountName": "Tmall","fullName": "","username": null,"profileUrl": "","url": "","description": "Brand cooperation and wholesale enquiries: brandcoop.export@gmail.com","email": "brandcoop.export@gmail.com","emailDomain": "@gmail.com","possiblyTruncated": false,"foundAt": "2026-08-31T09:42:17Z"}
Note the empty profileUrl and the null username. When Google prints no name.tmall.com host, there is no handle to recover and those fields stay blank by design.
Export the Tmall Email Scraper dataset as JSON, CSV, XLSX or HTML, or pull it through the Apify API into your CRM.
How to Use the Tmall Email Scraper
Step 1 — open the Actor. Launch the Tmall Email Scraper on the Apify platform.
Step 2 — set a small maxEmails. Something like 20 to 50. There is no point paying for a long run against a thin index.
Step 3 — use specific brand or category keywords. A named brand you already know sells on Tmall gives you a better chance than a generic term.
Step 4 — widen your email domains. Add @outlook.com, @hotmail.com, @163.com and @qq.com alongside the defaults; Chinese sellers rarely use Gmail.
Step 5 — read the run summary. It reports pages fetched, blocked pages and emails per page, so you can see immediately whether Google returned anything at all.
If the summary shows pages fetched but no results, that is the expected outcome documented throughout this page — not a misconfiguration.
Use Cases for the Tmall Email Scraper
| Use case | How the Tmall Email Scraper helps | Realistic expectation |
|---|---|---|
| Named-brand contact lookup | Search a specific brand plus cooperation or business | Best-case scenario for this Actor |
| Cross-border sourcing research | Confirm whether a Tmall Global page carries a public address | Occasional hits |
| Coverage completeness | Add Tmall to a multi-platform sweep so the gap is documented | Reliable, but usually returns little |
| Index auditing | Measure how much of tmall.com Google actually holds for your niche | Works as intended |
| CRM enrichment | Match any recovered shop label to accounts you already track | Very small volume |
| China market research | Read description snippets from whatever Google does surface | Thin sample |
Notice that every row above is framed modestly. Anyone promising a large Tmall seller email list from Google search is not describing something that exists.
When to Use the Tmall Email Scraper — and When to Use a Sibling
Use the Tmall Email Scraper when you specifically need Tmall, accept the volume, and want the search done properly rather than by hand.
Use a sibling when you need contactable Chinese suppliers at scale. The strongest options in this family are:
- DHgate Email Scraper — wholesale suppliers whose export pages Google indexes well.
- AliExpress Email Scraper — cross-border sellers with international-facing storefronts.
- Temu Email Scraper — the newer export-first marketplace from the same manufacturing base.
If your real target is Western storefronts rather than Chinese marketplaces, the footprint-based Shopify Store Email Scraper and Wix Stores Email Scraper return merchant domains at a far higher rate.
The Taobao Email Scraper shares Tmall's indexing problem, since it sits on the same side of the Great Firewall.
Limitations of the Tmall Email Scraper
These are real constraints. The first one dominates everything else.
- Very low volume by design of the index, not the Actor. Tmall is largely excluded from Google's index. A live test produced a single result block for the entire run. Baidu, not Google, is where these pages are properly indexed, and the Tmall Email Scraper does not read Baidu.
- Only publicly indexed emails. If an address is not visible in Google's index, it cannot be found. Private data is never accessed.
- Google's ~300-result cap. A single query returns roughly 300 results at most. Query expansion exists to work around this, though on Tmall the ceiling is rarely the binding constraint.
possiblyTruncated. When Google's snippet ellipsis touches an address, this flag is set totrue. Verify those rows before sending anything.- Apify GOOGLE_SERP proxy required. The Actor cannot run without Apify proxy credentials.
- Free-plan cap. Free Apify plans are limited to 100 emails per run. Paid plans are uncapped — a limit you are very unlikely to reach here.
usernameandprofileUrlavailability. These are only populated when Google's result exposes aname.tmall.comhandle. Most Tmall rows show only a display label, leavingusernameasnullandprofileUrlempty. That is a Google limitation, not a bug.- Language and domain mismatch. Tmall sellers overwhelmingly use Chinese mailbox providers, so the default
@gmail.comand@yahoo.comfilters suppress most of what little exists. - No guaranteed volume. Results vary with keywords, domains and location, and on this platform the variance floor is zero.
Nothing here is affiliated with, endorsed by or officially supported by Tmall or Alibaba Group.
Related Actors
| Actor | What it collects |
|---|---|
| Tmall Email and Phone Number Scraper | Emails and phone numbers from Tmall |
| Tmall Phone Number Scraper | Public phone numbers from Tmall |
| AliExpress Email Scraper | Public contact emails from AliExpress |
| Allegro Email Scraper | Public contact emails from Allegro |
| Amazon Email Scraper | Public contact emails from Amazon |
| Best Buy Seller Email Scraper | Public contact emails from Best Buy |
| BigCommerce Store Email Scraper | Public contact emails from BigCommerce Store |
| Cdiscount Email Scraper | Public contact emails from Cdiscount |
| Coupang Email Scraper | Public contact emails from Coupang |
| Depop Email Scraper | Public contact emails from Depop |
| DHgate Email Scraper | Public contact emails from DHgate |
| eBay Email Scraper | Public contact emails from eBay |
| Ecwid Store Email Scraper | Public contact emails from Ecwid Store |
| Etsy Email Scraper | Public contact emails from Etsy |
| Faire Email Scraper | Public contact emails from Faire |
| Flipkart Email Scraper | Public contact emails from Flipkart |
| Home Depot Seller Email Scraper | Public contact emails from Home Depot |
| JD.com Email Scraper | Public contact emails from JD.com |
| Lazada Email Scraper | Public contact emails from Lazada |
| Lowe's Seller Email Scraper | Public contact emails from Lowe's |
| MercadoLibre Email Scraper | Public contact emails from MercadoLibre |
| Mercari Email Scraper | Public contact emails from Mercari |
| Newegg Email Scraper | Public contact emails from Newegg |
| Otto Email Scraper | Public contact emails from Otto |
| Overstock Email Scraper | Public contact emails from Overstock |
| Poshmark Email Scraper | Public contact emails from Poshmark |
| Rakuten Email Scraper | Public contact emails from Rakuten |
| Shopee Email Scraper | Public contact emails from Shopee |
| Shopify Store Email Scraper | Public contact emails from Shopify Store |
| Taobao Email Scraper | Public contact emails from Taobao |
| Target Seller Email Scraper | Public contact emails from Target |
| Temu Email Scraper | Public contact emails from Temu |
| Vinted Email Scraper | Public contact emails from Vinted |
| Walmart Email Scraper | Public contact emails from Walmart |
| Wayfair Email Scraper | Public contact emails from Wayfair |
| Wish Email Scraper | Public contact emails from Wish |
| Wix Stores Email Scraper | Public contact emails from Wix Stores |
| WooCommerce Email Scraper | Public contact emails from WooCommerce |
| Zalando Email Scraper | Public contact emails from Zalando |
| AliExpress Email and Phone Number Scraper | Emails and phone numbers from AliExpress |
| Allegro Email and Phone Number Scraper | Emails and phone numbers from Allegro |
| Amazon Email and Phone Number Scraper | Emails and phone numbers from Amazon |
Tmall Email Scraper Example Run
A sourcing consultant is compiling a China supplier map and wants Tmall represented rather than silently skipped.
They run the Tmall Email Scraper with keywords: ["brand", "supplier"], customDomains: ["@gmail.com", "@outlook.com"] and maxEmails: 30, expecting little.
The run finishes quickly. The summary logs pages fetched and no blocked pages, confirming Google was reached and simply had almost nothing to return.
They record the gap in their notes, then re-run the same keywords through the DHgate Email Scraper, which is where the usable supplier contacts actually come from.
Tmall Email Scraper FAQ
How many results should I expect from the Tmall Email Scraper?
Very few, and often none. In a measured live test the run surfaced a single result block. Plan for a handful of rows at best, and do not build a campaign around this Actor.
Why is the yield so low — is something broken?
No. Tmall is largely excluded from Google's index. Baidu is the dominant search engine for Chinese marketplaces and is where these pages are properly crawled. The Tmall Email Scraper reads Google only.
Will more keywords or more pages fix it?
No. Expansion widens the search over the same thin corpus. It is worth leaving on, but it cannot create index coverage that does not exist.
Which sibling should I use instead?
For Chinese suppliers, the DHgate, AliExpress and Temu Actors. For Western storefronts, the footprint-based Shopify, Wix and Squarespace Actors return merchant domains at a far higher rate.
Does the Tmall Email Scraper log into Tmall?
No. It never logs in, never uses a Tmall API and never opens the site. All data comes from publicly indexed Google search results.
Why are username and profileUrl usually empty?
Because Google rarely exposes a name.tmall.com host in its result blocks. Without that host there is no handle to recover, so username stays null and profileUrl stays empty.
Should I change the email domains?
Yes. Tmall sellers overwhelmingly use Chinese providers, so adding domains such as @163.com and @qq.com gives the Tmall Email Scraper more to match against than Gmail alone.
Do I need an Apify proxy?
Yes. The Tmall Email Scraper requires the Apify GOOGLE_SERP proxy and cannot run without Apify proxy credentials.
How many emails can it collect per run?
Free Apify plans are capped at 100 emails per run and paid plans are uncapped — but on this platform the index, not the plan, is your ceiling.
What does possiblyTruncated: true mean?
Google's snippet ellipsis touched the email, so the address may be cut off. Verify those rows before sending.
What happens if a run is interrupted?
Progress is saved in the key-value store, keyed by a hash of your input, and flushed on PERSIST_STATE, MIGRATING and ABORTING. Restarting resumes rather than starting over.
Is this affiliated with Tmall?
No. The Tmall Email Scraper is an independent tool and is not affiliated with, endorsed by or officially supported by Tmall or Alibaba Group.
Responsible Use and GDPR
Publicly visible does not mean consent to bulk marketing. Treat every address the Tmall Email Scraper returns as personal data belonging to a real business contact.
Follow GDPR, CAN-SPAM, PIPL and any other regime that applies to you: identify yourself, state why you are contacting them, honour opt-outs immediately and delete records on request.
With volumes this low, individual, well-researched outreach is the only sensible approach anyway.
Leave a review
If the Tmall Email Scraper saved you time, please leave a star rating and a short review on the Actor page.
Reviews are how other buyers judge whether a tool works, and they tell us which features to build next.
If something did not work, email neurodata.apify@gmail.com instead - bugs get fixed faster than they get complained about.
Support
Questions, bug reports or a custom build request? Email neurodata.apify@gmail.com and include your run ID so the Tmall Email Scraper logs can be checked quickly.