# Taobao Products Scraper 淘宝 天猫: Best Sellers, Prices, Shops (`themineworks/taobao-products-scraper`) Actor

Scrape Taobao and Tmall (淘宝 天猫) best sellers by keyword or category: full title, price in CNY, promo price, shop, category, image and selling points. Add item details (ships from, seller scores, shop age) or look items up by id. No login, no cookies. Pay per product.

- **URL**: https://apify.com/themineworks/taobao-products-scraper.md
- **Developed by:** [The Mine Works](https://apify.com/themineworks) (community)
- **Categories:** E-commerce, MCP servers
- **Stats:** 1 total users, 0 monthly users, 0.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $1.99 / 1,000 products

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Taobao Products Scraper 淘宝 天猫: Best Sellers, Prices, Shops

[![182 Taobao products in 16 seconds](https://api.apify.com/v2/key-value-stores/cUXz95yxflDho41nn/records/taobao-products-scraper-hero-fix0930.png)](https://console.apify.com/actors/tl3xHG9gCeJf1Z5w4/input)

From **The Mine Works**, makers of [Threads Scraper](https://apify.com/themineworks/threads-scraper) and [B2B Leads Finder](https://apify.com/themineworks/b2b-leads-finder), with nearly 139,000 runs across our actors.

Give it Chinese product words such as 手机壳 or 机械键盘 and it returns the products Taobao ranks as best sellers and top rated for each word: the full title, the price in yuan, the promotion price, the shop, the category path, the main image and the seller's selling points. Turn on item details and each product also gets where it ships from, whether it is a Tmall store, the seller's three service scores, the shop's open date and size, and a recent review. You can also look up products by item id or by any Taobao or Tmall link. Everything is read from the pages Taobao publishes for every visitor at www.taobao.com/list/, the part of the site its robots.txt opens to crawlers, so there is no login, no cookie and no browser.

### Why choose this actor?

- **182 products from 6 keywords in 16 seconds.** A local test run on 1 October 2026 read 5 keywords plus one Taobao has no page for, made 13 page requests and returned 182 different products with prices, shops and categories. No Taobao account, no cookies, no app.
- **The products Taobao itself ranks.** Each keyword gives Taobao's 20 best sellers and up to 40 top rated products, about half of them new, with their rank and the total Taobao counts for the word. You see what sells, not a random page of listings.
- **You pay only for new products.** A product found again under a second keyword, a product outside your price range, a product seen in an earlier run with monitor mode on, and a word Taobao has no page for are never charged.

[![Run it on Apify](https://api.apify.com/v2/key-value-stores/cUXz95yxflDho41nn/records/button-run.png)](https://console.apify.com/actors/tl3xHG9gCeJf1Z5w4/input)

**Part of The Mine Works More tools family:** [G2 Reviews Scraper](https://apify.com/themineworks/g2-reviews-scraper), [Tennis Match & Player Data Scraper](https://apify.com/themineworks/tennis-match-data), [Flashscore Tennis Scraper](https://apify.com/themineworks/flashscore-tennis-results-scraper), [LandWatch Scraper](https://apify.com/themineworks/landwatch-land-for-sale-scraper), [Capterra Reviews Scraper](https://apify.com/themineworks/capterra-software-reviews-scraper), [Google Hotels Prices Scraper](https://apify.com/themineworks/google-hotels-prices-scraper).

### Try it in one minute

Paste this into the JSON tab of the input form and press Start. It returns about 40 products in under 15 seconds.

```json
{
  "keywords": ["手机壳"],
  "maxProductsPerKeyword": 40
}
```

You can give the actor three kinds of input, in any mix: **keywords** (a Chinese product word such as 连衣裙, or a `www.taobao.com/list/product/` link), **categories** (a Taobao category id such as `150704`, or a `www.taobao.com/list/category/` link) and **items** (an item id such as `971278517006`, or any Taobao or Tmall item link that carries the id, such as `https://item.taobao.com/item.htm?id=971278517006` or `https://detail.tmall.com/item.htm?id=1078907312106`).

Apify's free plan includes $5 of credit every month, which covers about 1,600 products at this actor's Free plan price of $2.99 per 1,000 products, plus a flat $0.005 per run whatever memory you choose.

#### Copy to your AI assistant

Paste this block into ChatGPT, Claude, Cursor or any assistant that can write code, and it can run the actor for you.

```
themineworks/taobao-products-scraper on Apify. Returns Taobao and Tmall best sellers and top rated products for Chinese product keywords or category ids (title, price_cny, promo_price_cny, shop_name, category_path, image_url, rank), and item details by item id or Taobao/Tmall link (ships_from, platform, seller scores dsr_description/dsr_service/dsr_logistics, shop_opened_at, shop_item_count, latest_review). Call ApifyClient("TOKEN").actor("themineworks/taobao-products-scraper").call(run_input={...}), then client.dataset(run["defaultDatasetId"]).list_items().items. Required: at least one of keywords (array of Chinese product words or www.taobao.com/list/product/ links), categories (array of category ids), items (array of item ids or item links). Optional: includeTopRated (boolean, default true), addItemDetails (boolean, default false), maxProductsPerKeyword (integer 1 to 60, default 60), maxProducts (integer, default 1000), minPrice and maxPrice (yuan), onlyNewProducts (boolean, default false). Rows with _type "info" explain a run that delivered nothing. Full spec: GET https://api.apify.com/v2/acts/themineworks~taobao-products-scraper/builds/default (Bearer TOKEN), which returns inputSchema and readme. Token: https://console.apify.com/account/integrations
```

### Key features

- **About 40 different products per keyword.** Taobao's best seller page gives 20 products and its top rated page up to 40. On our test keywords 20 to 22 of the top rated products were not among the best sellers, so a keyword yields about 40 different products. A few words have a best seller page but no top rated page; they give 20.
- **About 30 fields per product, 50 with details.** Title, item id, price in yuan, promotion price and its name, prices Taobao converts to HKD and TWD, shop id and name, leaf and top level category ids with the full category path, image, selling points, the buyers' description score, free shipping, rank and the keyword's total on Taobao.
- **Item details on demand.** With `addItemDetails` on, the actor opens each product's item page and adds 20 fields: Tmall or Taobao, ships from, the price on the item page, the seller's description, service and logistics scores, seller type and credit level, shop open date, shop size, recent additions, payment methods, a recent review with the variant bought, and Taobao's own summary of what buyers say.
- **Item lookups by id or link.** Paste up to 1,000 item ids or Taobao and Tmall links. The actor reads the id from the link and never opens item.taobao.com or detail.tmall.com, whose robots.txt closes them to crawlers.
- **Price filter and run caps.** `minPrice` and `maxPrice` in yuan, `maxProductsPerKeyword` and `maxProducts` keep a run to what you need, and anything left out is never charged.
- **Monitor mode for schedules.** With `onlyNewProducts` on, the actor remembers up to 50,000 products this input delivered and returns only new ones on later runs, so a daily schedule shows you new best sellers.

### How to use it

#### Basic: best sellers for one word

```json
{
  "keywords": ["蓝牙耳机"]
}
```

Returns up to 60 products: the 20 best sellers first (`source_list` is `best-selling`), then the top rated ones that were not already delivered (`top-rated`). Each row carries its `rank` on the Taobao list.

#### Several words at once, best sellers only

```json
{
  "keywords": ["连衣裙", "保温杯", "机械键盘", "蓝牙耳机"],
  "includeTopRated": false
}
```

Four words, 20 products each, in a few seconds. A product that shows up under two words is delivered once, under the first word, and charged once.

#### Market research: a category's best sellers with seller scores

```json
{
  "categories": ["150704", "50010850"],
  "addItemDetails": true
}
```

Category ids are in every row (`category_id`, `root_category_id`), so you can take them from an earlier keyword run. With `addItemDetails` on, every product also gets `ships_from`, `platform` (tmall or taobao), `dsr_description`, `dsr_service`, `dsr_logistics`, `shop_opened_at` and `shop_item_count`, which is enough to tell a twelve year old flagship store from a new reseller.

#### Sourcing: check a list of supplier links

```json
{
  "items": [
    "https://item.taobao.com/item.htm?id=971278517006",
    "https://detail.tmall.com/item.htm?id=1078907312106",
    "991372610509"
  ]
}
```

One row per item with the price on the item page, where it ships from, the shop and its scores, and a recent review with the variant the buyer chose. Taobao shortens the title on this page (`title_is_complete` is false); run the product's keyword to get the full title.

#### Daily price watch on a price band

```json
{
  "keywords": ["手机壳", "保温杯"],
  "minPrice": 20,
  "maxPrice": 80,
  "onlyNewProducts": true
}
```

Schedule it daily in Apify Console (Schedules, Add schedule). Each run delivers only products that entered the best seller or top rated lists since the last run, inside 20 to 80 yuan. Products seen before and products outside the band are never charged.

### Input parameters

| Parameter | Type | Default | What it does |
|---|---|---|---|
| `keywords` | array of strings | none | Product words, best in Chinese (手机壳, 连衣裙), or `www.taobao.com/list/product/` links. Each gives Taobao's 20 best sellers and, with `includeTopRated`, up to 40 top rated products. Up to 200 per run. |
| `categories` | array of strings | none | Taobao category ids (150704) or `www.taobao.com/list/category/` links. Each gives the category's 20 best sellers. Up to 100 per run. |
| `items` | array of strings | none | Item ids or Taobao and Tmall item links carrying the id. One row each with item page details. Up to 1,000 per run. |
| `includeTopRated` | boolean | `true` | Also read each keyword's top rated list. |
| `addItemDetails` | boolean | `false` | Open each listed product's item page and add 20 detail fields. One more page per product, charged as item details (see Pricing) only when the item page was read. |
| `maxProductsPerKeyword` | integer | `60` | Most products delivered per keyword or category, 1 to 60. |
| `maxProducts` | integer | `1000` | The run stops after this many products, 1 to 20,000. |
| `minPrice` | integer | none | Leave out products below this price in yuan (promo price when there is one). Never charged. |
| `maxPrice` | integer | none | Leave out products above this price in yuan. Never charged. |
| `onlyNewProducts` | boolean | `false` | Deliver only products this input has not delivered before. Kept per set of keywords, categories and items. |

At least one of `keywords`, `categories` or `items` is needed. A run with none of them stops at once, writes one info row and charges nothing, not even the start fee.

### What data do you get?

One row per product. Fields are grouped below; empty fields are left out of a row rather than sent as null.

- **Product:** `item_id`, `title`, `title_is_complete`, `url` (the Taobao item link, which opens the Tmall page for Tmall items), `taobao_list_url`, `image_url`, `images`, `selling_points`, `description_match_score`, `free_shipping`, `area_limited`, `review_tags`.
- **Price:** `price_cny`, `promo_price_cny`, `promo_label`, `price_hkd`, `price_twd`, `currency`.
- **Shop and category:** `shop_id`, `shop_name`, `category_id`, `root_category_id`, `category_path`.
- **Where it came from:** `source` (keyword, category or item), `source_list` (best-selling or top-rated), `keyword`, `input_category_id`, `rank`, `list_total_count`, `input`, `scraped_at`.
- **Item page details** (item inputs, and every row when `addItemDetails` is on; `details_added` says which): `platform`, `tmall_url`, `ships_from`, `page_price_cny`, `page_original_price_cny`, `has_price_range`, `seller_user_id`, `seller_type`, `seller_credit_level`, `shop_url`, `shop_opened_at`, `shop_item_count`, `shop_new_item_count`, `dsr_description`, `dsr_service`, `dsr_logistics`, `payment_methods`, `latest_review` (rating, text, date, variant), `has_more_reviews`, `review_summary`, and `details_error` when the item page could not be read.

#### Stable fields for automations

These fields were present in every one of the 182 rows of the test run, and in every item row. Their names will not change.

| Field | What it holds |
|---|---|
| `item_id` | Taobao item id, as text |
| `title` | Product title (full on keyword and category rows) |
| `title_is_complete` | false when Taobao shortened the title |
| `url` | Taobao item link |
| `price_cny` | Price in yuan |
| `currency` | Always CNY |
| `shop_id` | Taobao shop id |
| `shop_name` | Shop name |
| `category_id` | Leaf category id |
| `root_category_id` | Top level category id |
| `category_path` | Category breadcrumb, top level first |
| `image_url` | Main image |
| `source` | keyword, category or item |
| `details_added` | true when the item page details are filled |
| `scraped_at` | When the row was saved, ISO 8601 |

#### Output examples

A best seller for 连衣裙 (dress), without item details:

```json
{
  "item_id": "1079130401408",
  "title": "DoggyQin 蚂蚁腰/粗棒高克重绵羊毛连衣裙女法式毛衣外套毛茸裙裤",
  "title_is_complete": true,
  "url": "https://item.taobao.com/item.htm?id=1079130401408",
  "price_cny": 469,
  "promo_price_cny": 469,
  "promo_label": "超级立减活动价",
  "price_hkd": 554.17,
  "shop_id": "114424207",
  "shop_name": "DoggyQin",
  "category_id": "50010850",
  "root_category_id": "16",
  "category_path": "女装 > 连衣裙 > 连衣裙",
  "image_url": "https://img.alicdn.com/imgextra/i2/2275483160/O1CN01huUiao0q2IJ2vH2e_!!2275483160.jpg",
  "description_match_score": 4.78,
  "free_shipping": false,
  "source": "keyword",
  "source_list": "best-selling",
  "keyword": "连衣裙",
  "rank": 1,
  "list_total_count": 66545,
  "details_added": false
}
```

A top rated product for 保温杯 (vacuum flask), with the seller's selling points:

```json
{
  "item_id": "826563498695",
  "title": "迪士尼迷你保温杯儿童水杯小巧水壶便携幼儿园宝宝吸管口袋杯子",
  "price_cny": 79,
  "promo_price_cny": 79,
  "promo_label": "新品秒杀",
  "shop_name": "Disney迪士尼文具旗舰店",
  "category_path": "餐饮具 > 杯子/水壶 > 保温杯 > 保温杯",
  "selling_points": ["迷你杯身 便携实用", "食品硅胶吸管 安全健康"],
  "description_match_score": 4.86,
  "source": "keyword",
  "source_list": "top-rated",
  "keyword": "保温杯",
  "rank": 10,
  "list_total_count": 32471,
  "details_added": false
}
```

A best seller for 机械键盘 (mechanical keyboard) with `addItemDetails` on. The item page price can differ from the list price while a promotion runs:

```json
{
  "item_id": "991372610509",
  "title": "机械手感有线静音电脑键盘鼠标套装笔记本游戏电竞外设三件套办公",
  "price_cny": 34.8,
  "shop_name": "天天特卖工厂",
  "category_path": "电脑硬件/显示器/电脑周边 > 键盘 > 机械键盘",
  "source_list": "best-selling",
  "keyword": "机械键盘",
  "rank": 1,
  "platform": "tmall",
  "tmall_url": "https://detail.tmall.com/item.htm?id=991372610509",
  "ships_from": "江西宜春",
  "page_price_cny": 28.8,
  "seller_type": "tmall_or_business",
  "seller_credit_level": 20,
  "shop_opened_at": "2018-05-11 18:37:50",
  "shop_item_count": 122302,
  "dsr_description": 4.7,
  "dsr_service": 4.8,
  "dsr_logistics": 4.8,
  "payment_methods": ["蚂蚁花呗", "信用卡支付", "集分宝"],
  "latest_review": {
    "rating": 5,
    "text": "键盘膜非常好，手感非常好，外表好看，而且价格便宜...",
    "date": "2026.09.22 09:50",
    "variant": "颜色分类:黑色蓝光【单键盘】"
  },
  "has_more_reviews": true,
  "details_added": true
}
```

An item looked up from a Tmall link (`https://detail.tmall.com/item.htm?id=1078907312106`):

```json
{
  "item_id": "1078907312106",
  "title": "【新款18勃艮第红色...",
  "title_is_complete": false,
  "price_cny": 116,
  "shop_name": "图拉斯数码旗舰店",
  "category_path": "阿里B2C商城 > 手机及配件 > 手机配件 > 保护套/硅胶套",
  "source": "item",
  "platform": "tmall",
  "ships_from": "广东东莞",
  "shop_opened_at": "2012-05-28 11:41:20",
  "shop_item_count": 242,
  "dsr_description": 4.7,
  "dsr_service": 4.8,
  "dsr_logistics": 4.8,
  "latest_review": {
    "rating": 5,
    "date": "2026.09.18 16:28",
    "variant": "适用手机型号:iPhone 18 Pro Max;颜色分类:【雪雾 | 焕新紫 | 磁吸款】冰感磁吸✅️细腻抗指纹✅️超薄裸感"
  },
  "details_added": true
}
```

Reviewer names are masked by Taobao itself and are not collected.

### Pricing

Pay per event: you pay for products delivered, plus a flat start fee.

| Event | Free | Bronze | Silver | Gold and above |
|---|---|---|---|---|
| Product delivered (per 1,000) | $2.99 | $2.69 | $2.39 | $1.99 |
| Item details added, with `addItemDetails` on (per 1,000) | $3.00 | $2.70 | $2.40 | $2.00 |
| Run start (once per run) | $0.005 | $0.005 | $0.005 | $0.005 |

With `addItemDetails` on, each product whose item page was read also pays the item details fee, so a product with details costs $5.99 per 1,000 on the Free plan ($3.99 on Business). A product whose item page could not be read is charged as a plain product. Plus a flat $0.005 per run whatever memory you choose.

Never charged:

- a product already delivered in the same run (found under a second keyword or in both lists);
- products outside `minPrice` and `maxPrice`;
- products delivered in an earlier run when `onlyNewProducts` is on;
- keywords and categories Taobao has no page for, and items Taobao no longer has;
- pages Taobao refused;
- the info rows (`_type: "info"`): the note that explains a run that delivered nothing, and the closing note at the end of a run that did;
- a run whose input has nothing to scrape (not even the start fee).

You can cap spending in Apify Console with the run's maximum charge; the actor stops cleanly when it is reached.

### FAQ

#### What is Taobao, and does this cover Tmall too?

Taobao (淘宝) is Alibaba's largest consumer marketplace in China, and Tmall (天猫) is its section for brand and flagship stores. Both share one catalogue, so a keyword returns Taobao and Tmall products together. With item details on, `platform` says which one a product belongs to and `tmall_url` gives the Tmall link.

#### How many products can I get per keyword?

At most 60, and about 40 different ones in practice: Taobao shows visitors 20 best sellers and up to 40 top rated products per word, and the two lists overlap. Taobao's further pages are not open to visitors (asking for page 2 sends you to world.taobao.com), so for more products, give more words. Narrower words (蓝牙耳机 rather than 耳机) surface different products. `list_total_count` tells you how many products Taobao counts for the word.

#### Which keywords work?

Common product words in Chinese work best, because Taobao builds a page for each word it sees searched often. In our test 5 of 6 words had a page; "mechanical keyboard" in English had none, while 机械键盘 returned 40 products. A word without a page is reported in the run summary as `no_taobao_page` and costs nothing. A link to a `www.taobao.com/list/product/` page always names a word that has a page.

#### Can I get product details from an item link?

Yes. Paste item ids or links into `items`. You get the item page's price, ships from, the shop with its scores, size and open date, payment methods and a recent review. Taobao shortens titles on this page, so `title_is_complete` is false there; the same product from a keyword or category run has its full title. Variant (SKU) prices, stock, full image galleries and full review lists are not on the pages this actor reads.

#### Do I need a Taobao account, cookies or an app?

No. The actor reads public pages as an anonymous visitor. It never logs in, never asks for cookies and never opens an app.

#### Why only www.taobao.com/list/ pages?

Because that is the part of Taobao its robots.txt opens to crawlers (`Allow: /list/*` for every user agent). The search site, item.taobao.com, detail.tmall.com and world.taobao.com all say `Disallow: /` to crawlers other than named search engines, so the actor never opens them, even when you paste their links (it only reads the id out of the link). The robots.txt file is read fresh at the start of every run, and every URL is checked against it, redirects included.

#### How fresh is the data?

Live. Every run reads Taobao's pages at that moment; nothing comes from a cache or database. Taobao's lists shift during the day as sales move, so a daily schedule with `onlyNewProducts` shows you what entered the lists.

#### What proxy does it use?

Apify's datacenter proxy, included in your plan. You do not need to set anything. In our tests every one of 38 page requests from datacenter addresses was answered on the first try, with no verification page. If Taobao ever answers with a verification page, the actor does not try to get past it: the page counts as refused, it tries again later on a new address, and it stops the run if refusals keep coming, without charging for anything not delivered.

#### What prices are in the rows?

`price_cny` is the price Taobao shows on the list, and `promo_price_cny` with `promo_label` the promotion price when one runs. The HKD and TWD prices are Taobao's own conversions. With item details, `page_price_cny` is the price on the item page at that moment, which can differ from the list price while a promotion runs. Prices are in yuan and before shipping.

#### Can I run it on a schedule and get only new products?

Yes. Create a schedule in Apify Console and turn on `onlyNewProducts`. The actor remembers up to 50,000 products per set of keywords, categories and items, and later runs deliver and charge only new ones.

#### What formats can I export?

JSON, CSV, Excel, XML, HTML and RSS from the dataset, or straight into Google Sheets, a webhook or the API.

#### Can I use it from Claude, ChatGPT or another AI assistant?

- Connector URL: `https://mcp.apify.com/?tools=themineworks/taobao-products-scraper`.
- Claude: Settings > Connectors > Add custom connector, paste the URL, sign in with Apify.
- ChatGPT: developer mode, add an MCP connector with the URL, sign in with Apify.
- Cursor or VS Code: add it as an HTTP MCP server with that URL.
- Claude Code: `claude mcp add -t http taobao-products-scraper "https://mcp.apify.com/?tools=themineworks/taobao-products-scraper"`.

#### Is it legal to scrape Taobao?

The actor reads only public product pages that Taobao's robots.txt opens to crawlers, without logging in. Product, price and shop data is business data. Reviewer names are masked by Taobao and not collected. You are responsible for how you use the data, including Taobao's terms and data protection laws such as GDPR, CCPA and China's PIPL.

### Integrations

- **Google Sheets:** send each run's dataset to a sheet with Apify's Google Sheets integration.
- **Make, Zapier and n8n:** start a run and pick up the products in your own flow.
- **Webhooks:** get a call when a run finishes, with the dataset link.
- **API and SDKs:** run it from Python or JavaScript with the Apify client, as in the block above.
- **MCP clients:** Claude, Cursor and other MCP clients can call it through mcp.apify.com.

### More from The Mine Works

**More tools**

- [G2 Reviews Scraper](https://apify.com/themineworks/g2-reviews-scraper)
- [Tennis Match & Player Data Scraper](https://apify.com/themineworks/tennis-match-data)
- [Flashscore Tennis Scraper](https://apify.com/themineworks/flashscore-tennis-results-scraper)
- [LandWatch Scraper](https://apify.com/themineworks/landwatch-land-for-sale-scraper)
- [Capterra Reviews Scraper](https://apify.com/themineworks/capterra-software-reviews-scraper)
- [Google Hotels Prices Scraper](https://apify.com/themineworks/google-hotels-prices-scraper)
- [Google Lens OCR Scraper](https://apify.com/themineworks/google-lens-ocr-scraper)
- [WeChat Article Scraper 微信公众号文章](https://apify.com/themineworks/wechat-article-scraper)

**Social media and video**

- [Threads Scraper](https://apify.com/themineworks/threads-scraper)
- [Reddit Scraper](https://apify.com/themineworks/reddit-scraper)
- [Threads Search Scraper](https://apify.com/themineworks/threads-search-scraper)
- [Instagram Profile Scraper](https://apify.com/themineworks/instagram-profile-scraper)

**Leads and business directories**

- [B2B Leads Finder](https://apify.com/themineworks/b2b-leads-finder)
- [Skip Trace Lookup](https://apify.com/themineworks/skip-trace-lookup)
- [Google Maps Email Scraper](https://apify.com/themineworks/maps-leads)
- [JustDial Scraper](https://apify.com/themineworks/justdial-business)

**Marketing, SEO and reviews**

- [Facebook Ad Library Scraper](https://apify.com/themineworks/meta-ad-library-scraper)
- [Google Ads Transparency Scraper](https://apify.com/themineworks/google-ads-transparency)
- [Similarweb Scraper](https://apify.com/themineworks/similarweb-scraper)
- [Google News Scraper](https://apify.com/themineworks/google-news)

**LinkedIn**

- [LinkedIn Company Scraper](https://apify.com/themineworks/linkedin-company-details)
- [LinkedIn Post Scraper](https://apify.com/themineworks/linkedin-post-search)
- [LinkedIn Employees Scraper](https://apify.com/themineworks/linkedin-employees)
- [LinkedIn Profile Scraper](https://apify.com/themineworks/linkedin-profile-scraper)

**Real estate**

- [Zillow Rentals Scraper](https://apify.com/themineworks/zillow-rental-listings)
- [Zillow Sold Comps Scraper](https://apify.com/themineworks/zillow-recently-sold)
- [Housing.com Scraper](https://apify.com/themineworks/housing-com-scraper)
- [India Real Estate MCP](https://apify.com/themineworks/india-real-estate-mcp)

**Science, health and government data**

- [CourtListener Scraper](https://apify.com/themineworks/courtlistener-court-records)
- [Socrata Open Data Scraper](https://apify.com/themineworks/socrata-open-data)
- [Academic Research MCP](https://apify.com/themineworks/academic-research-mcp)
- [OpenAlex Scraper](https://apify.com/themineworks/openalex-scholarly-works)

**Jobs and hiring**

- [Foundit Monster India Jobs](https://apify.com/themineworks/foundit-jobs-scraper)
- [Hirist Jobs Scraper](https://apify.com/themineworks/hirist-jobs-scraper)
- [India Jobs MCP](https://apify.com/themineworks/india-jobs-mcp)
- [Naukri Jobs Scraper](https://apify.com/themineworks/naukri-jobs)

**Company and business data**

- [GST Taxpayer Lookup](https://apify.com/themineworks/gst-taxpayer-lookup)
- [Company Domain Finder](https://apify.com/themineworks/company-domain-finder)
- [World Bank Trade Scraper](https://apify.com/themineworks/global-trade-data)
- [SEC EDGAR Filings Scraper](https://apify.com/themineworks/sec-edgar-filings)

**E-commerce and marketplaces**

- [Ozon.ru Scraper](https://apify.com/themineworks/ozon-product-search)
- [Amazon Product Scraper](https://apify.com/themineworks/amazon-products)
- [⭐ Amazon Reviews Scraper](https://apify.com/themineworks/amazon-reviews)
- [Carsales.com.au Scraper](https://apify.com/themineworks/carsales-scraper)

**Food and local services**

- [NoBroker Scraper](https://apify.com/themineworks/nobroker-scraper)
- [Swiggy Restaurant Scraper](https://apify.com/themineworks/swiggy-scraper)
- [Zomato Scraper](https://apify.com/themineworks/zomato-scraper)

**Developer and AI tools**

- [Website to Markdown Crawler](https://apify.com/themineworks/rag-crawler)
- [GitHub Repo Scraper](https://apify.com/themineworks/github-repo-intelligence)
- [GitHub Skill Finder](https://apify.com/themineworks/github-skill-discovery)
- [GitHub Trending Scraper](https://apify.com/themineworks/github-trending-scraper)

### Support

Found a problem or want a field added? Open an issue on the Issues tab and include the run link. For a new marketplace or data source, email dmineworks@gmail.com.

*Taobao Products Scraper returns the products Taobao and Tmall rank as best sellers and top rated for your keywords and categories, with prices, shops, categories and seller scores, read from the pages Taobao opens to every visitor.*

# Actor input Schema

## `keywords` (type: `array`):

One per line: a product word, best in Chinese (手机壳, 连衣裙, 机械键盘, 蓝牙耳机), or a www.taobao.com/list/product/ link. Each gives Taobao's 20 best sellers for the word, plus up to 40 top rated ones. Taobao only has pages for common product words; any other word is reported as having no page and costs nothing. Up to 200 per run.

## `categories` (type: `array`):

One per line: a Taobao category id (150704) or a www.taobao.com/list/category/ link. Each gives the category's 20 best sellers. Category ids are in every row (category\_id, root\_category\_id). Up to 100 per run.

## `items` (type: `array`):

One per line: a Taobao item id (971278517006) or any Taobao or Tmall item link that carries the id (item.taobao.com/item.htm?id=..., detail.tmall.com/item.htm?id=...). Each gives one row with price, ships from, the seller's three scores, shop age and size and a recent review. Taobao shortens the title on this page; keyword rows carry the full title. Short m.tb.cn links are not read. Up to 1,000 per run.

## `includeTopRated` (type: `boolean`):

For each keyword, also read Taobao's top rated list (up to 40 products, about half of them not among the 20 best sellers). Off: only the 20 best sellers.

## `addItemDetails` (type: `boolean`):

Open each listed product's item page and add where it ships from, Tmall or Taobao, the seller's description, service and logistics scores, shop open date, shop size, payment methods and a recent review. One more page per product, so the run takes longer; the price per product stays the same.

## `maxProductsPerKeyword` (type: `integer`):

Most products delivered for one keyword or category. Taobao shows at most 20 best sellers and 40 top rated per word, so 60 is the ceiling.

## `maxProducts` (type: `integer`):

The run stops once this many products are delivered.

## `minPrice` (type: `integer`):

Leave out products below this price in yuan (the promo price when there is one). Products left out are never charged.

## `maxPrice` (type: `integer`):

Leave out products above this price in yuan. Products left out are never charged.

## `onlyNewProducts` (type: `boolean`):

For schedules: remember every product this input delivered and deliver only new ones next time. Products seen before are skipped and never charged. The memory is kept per set of keywords, categories and items.

## Actor input object example

```json
{
  "keywords": [
    "手机壳"
  ],
  "includeTopRated": true,
  "addItemDetails": false,
  "maxProductsPerKeyword": 60,
  "maxProducts": 1000,
  "onlyNewProducts": false
}
```

# Actor output Schema

## `results` (type: `string`):

No description

## `summary` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "keywords": [
        "手机壳"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("themineworks/taobao-products-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "keywords": ["手机壳"] }

# Run the Actor and wait for it to finish
run = client.actor("themineworks/taobao-products-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "keywords": [
    "手机壳"
  ]
}' |
apify call themineworks/taobao-products-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,themineworks/taobao-products-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/tl3xHG9gCeJf1Z5w4/builds/KPC7TgM2rxqDtqs4H/openapi.json
