# WeChat Article Scraper 微信公众号文章: Search and Full Text (`themineworks/wechat-article-scraper`) Actor

Search WeChat Official Account articles (微信公众号文章) by keyword through Sogou WeChat search: title, snippet, account, publish time, cover. Paste mp.weixin.qq.com links for full text, account name, WeChat id, gh\_ id, publish time and region. No login, no cookies. Pay per row.

- **URL**: https://apify.com/themineworks/wechat-article-scraper.md
- **Developed by:** [The Mine Works](https://apify.com/themineworks) (community)
- **Categories:** Social media, News, MCP servers
- **Stats:** 1 total users, 0 monthly users, 0.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.49 / 1,000 article or search results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## WeChat Article Scraper 微信公众号文章: Search and Full Text

[![58 search results, 3 full articles in 24 s](https://api.apify.com/v2/key-value-stores/cUXz95yxflDho41nn/records/wechat-article-scraper-hero-fix0930.png)](https://console.apify.com/actors/6XA30YmC0puKLClUF/input)

From **The Mine Works**, makers of [Threads Scraper](https://apify.com/themineworks/threads-scraper) and [B2B Leads Finder](https://apify.com/themineworks/b2b-leads-finder), with nearly 139,000 runs across our actors.

WeChat Official Accounts (微信公众号) are where Chinese brands, government bodies, media and analysts publish. This actor does two things with them. Give it keywords and it searches WeChat articles through Sogou's WeChat search, the only public index of them, and returns each result's title, excerpt, account name, publish time and cover. Give it article links (mp.weixin.qq.com/s/...) and it returns the full text with the account's name, WeChat id, gh\_ id, the author, the publish time in UTC and China time, the region WeChat shows under the article, the cover and the image count. Plain HTTP, no WeChat account, no cookies, no browser.

### Why choose this actor?

- **61 rows in 24 seconds.** A local test run on 1 October 2026 searched two keywords (人工智能 and 新能源汽车) three pages deep and read three article links: 58 search results and 3 full articles, 11 page requests, none refused. No WeChat login, no QR code, no phone.
- **Full text you can feed to a model.** Article rows carry the whole text with paragraph breaks (up to 50,000 characters by default), plus `content_length`, `image_count` and the account's own description, so you can summarise, classify or translate without opening WeChat.
- **You pay for rows, not for misses.** A result that shows up twice, a result seen in an earlier run with monitor mode on, an article that was deleted or whose account moved, and a refused page are never charged.

[![Run it on Apify](https://api.apify.com/v2/key-value-stores/cUXz95yxflDho41nn/records/button-run.png)](https://console.apify.com/actors/6XA30YmC0puKLClUF/input)

**Part of The Mine Works More tools family:** [G2 Reviews Scraper](https://apify.com/themineworks/g2-reviews-scraper), [Tennis Match & Player Data Scraper](https://apify.com/themineworks/tennis-match-data), [Flashscore Tennis Scraper](https://apify.com/themineworks/flashscore-tennis-results-scraper), [LandWatch Scraper](https://apify.com/themineworks/landwatch-land-for-sale-scraper), [Capterra Reviews Scraper](https://apify.com/themineworks/capterra-software-reviews-scraper), [Google Hotels Prices Scraper](https://apify.com/themineworks/google-hotels-prices-scraper).

### Try it in one minute

Paste this into the JSON tab of the input form and press Start. It returns 10 search results and one full article in about 15 seconds.

```json
{
  "keywords": ["新能源汽车"],
  "maxResultsPerKeyword": 10,
  "articleUrls": ["https://mp.weixin.qq.com/s/FN4I4jDgFcyEsv7T2OFzCg"]
}
```

There are two ways to give input, in any mix: **keywords** (any words, best in Chinese, or a pasted `weixin.sogou.com` search link) and **article links** (`https://mp.weixin.qq.com/s/<id>` or the long form `https://mp.weixin.qq.com/s?__biz=...&mid=...&idx=...&sn=...`; tracking parameters are dropped).

Apify's free plan includes $5 of credit every month, which covers about 1,250 rows at this actor's Free plan price of $3.99 per 1,000 rows, plus a flat $0.005 per run whatever memory you choose.

#### Copy to your AI assistant

Paste this block into ChatGPT, Claude, Cursor or any assistant that can write code, and it can run the actor for you.

```
themineworks/wechat-article-scraper on Apify. Searches WeChat Official Account articles (微信公众号文章) by keyword through Sogou WeChat search (rows with type "search_result": title, snippet, account_name, published_at, cover_image_url, sogou_link) and reads public article links (rows with type "article": title, content_text, account_name, account_wechat_id, account_gh_id, author, published_at, publisher_region, image_count, is_original). Call ApifyClient("TOKEN").actor("themineworks/wechat-article-scraper").call(run_input={...}), then client.dataset(run["defaultDatasetId"]).list_items().items. Required: at least one of keywords (array of strings) or articleUrls (array of mp.weixin.qq.com/s links). Optional: maxResultsPerKeyword (integer 1 to 100, default 30), includeFullText (boolean, default true), maxTextLength (integer, default 50000), maxItems (integer, default 1000), onlyNewArticles (boolean, default false). Search rows do not carry the mp.weixin.qq.com link. Rows with _type "info" explain a run that delivered nothing. Full spec: GET https://api.apify.com/v2/acts/themineworks~wechat-article-scraper/builds/default (Bearer TOKEN), which returns inputSchema and readme. Token: https://console.apify.com/account/integrations
```

### Key features

- **Up to 100 search results per keyword.** Sogou shows 10 results a page and at most 10 pages to a visitor; the actor reads them in order, with `page`, `position` and Sogou's estimate of total results on every row.
- **13 fields per search result, 30 per article.** Search rows: title, excerpt, account, publish time, cover, Sogou link and document id. Article rows: title, full text, length, account name, WeChat id, gh\_ id, biz id, avatar, account description, author, summary, cover, publish time in UTC and China time, region, country, message ids, the Original badge, the read more link and image and video counts.
- **Full text, cleaned.** Tags, styles and image placeholders are removed and paragraphs kept, so `content_text` reads like the article. `maxTextLength` cuts very long reports and says so in `content_truncated`.
- **Deleted and moved articles reported, never charged.** An article WeChat no longer shows (deleted by the publisher, removed for a rule, account moved, link expired) appears in the run summary with the reason and costs nothing.
- **Monitor mode for schedules.** With `onlyNewArticles` on, the actor remembers up to 50,000 results and articles per input and returns only new ones on later runs.

### How to use it

#### Basic: search one topic

```json
{
  "keywords": ["人工智能"]
}
```

Returns up to 30 results (3 Sogou pages), newest and most relevant first as Sogou orders them, each with `page` and `position`.

#### Several topics, as deep as Sogou goes

```json
{
  "keywords": ["新能源汽车", "储能", "固态电池"],
  "maxResultsPerKeyword": 100
}
```

Up to 10 pages per keyword. A result that appears under two keywords is delivered once, under the first, and charged once.

#### Full text of articles you already have

```json
{
  "articleUrls": [
    "https://mp.weixin.qq.com/s/FN4I4jDgFcyEsv7T2OFzCg",
    "https://mp.weixin.qq.com/s/f9uC5ztnJvNy5jaxQL2xQg",
    "https://mp.weixin.qq.com/s/-elVpDBa4F2w2MKYb7iUrA"
  ],
  "maxTextLength": 20000
}
```

One row per article with the text and the account behind it. Links shared in WeChat chats, on Weibo or in newsletters all work, short or long form.

#### Brand monitoring: new mentions every morning

```json
{
  "keywords": ["比亚迪", "蔚来"],
  "maxResultsPerKeyword": 20,
  "onlyNewArticles": true
}
```

Schedule it daily in Apify Console (Schedules, Add schedule). Each run delivers only results that were not delivered before, so the dataset is your list of new WeChat mentions. Results seen before are never charged.

#### Search rows without the text, for a quick scan

```json
{
  "keywords": ["跨境电商"],
  "maxResultsPerKeyword": 50,
  "includeFullText": false
}
```

`includeFullText` only affects article rows; search rows never carry the text.

### Input parameters

| Parameter | Type | Default | What it does |
|---|---|---|---|
| `keywords` | array of strings | none | Words to search WeChat articles for through Sogou's WeChat search, or `weixin.sogou.com` search links. Up to 100 per run. |
| `articleUrls` | array of strings | none | Public article links, `mp.weixin.qq.com/s/<id>` or `mp.weixin.qq.com/s?__biz=...&mid=...&idx=...&sn=...`. One row each with full text. Up to 500 per run. |
| `maxResultsPerKeyword` | integer | `30` | Most search results per keyword, 1 to 100 (10 per Sogou page, 10 pages at most). |
| `includeFullText` | boolean | `true` | Article rows carry the full text in `content_text`. Off: every other field, smaller rows, same price. |
| `maxTextLength` | integer | `50000` | Longest text kept per article, 200 to 200,000 characters; longer text is cut and `content_truncated` is true. |
| `maxItems` | integer | `1000` | The run stops after this many rows, 1 to 10,000. |
| `onlyNewArticles` | boolean | `false` | Deliver only results and articles this input has not delivered before. |

At least one of `keywords` or `articleUrls` is needed. A run with neither stops at once, writes one info row and charges nothing, not even the start fee.

### What data do you get?

Every row has a `type`: `search_result` (from a keyword) or `article` (from a link). Empty fields are left out of a row rather than sent as null.

- **Search results:** `keyword`, `page`, `position`, `title`, `snippet`, `account_name`, `published_at`, `cover_image_url`, `sogou_link`, `sogou_doc_id`, `total_results_estimate`.
- **Articles, the text:** `title`, `content_text`, `content_length`, `content_truncated`, `description`, `image_count`, `video_count`, `item_show_type`, `is_original`, `copyright_stat`, `source_url`.
- **Articles, the account:** `account_name`, `account_wechat_id`, `account_gh_id`, `account_biz`, `account_avatar_url`, `account_signature`, `author`.
- **Articles, when and where:** `published_at` (UTC), `published_at_china_time`, `publisher_region`, `publisher_country`, `url`, `canonical_url`, `mid`, `idx`, `sn`, `cover_image_url`.
- **Every row:** `type`, `input`, `scraped_at`.

Search rows carry Sogou's link (`sogou_link`), not the mp.weixin.qq.com link: Sogou only reveals the article link to a browser click, and an automated request for it gets Sogou's verification page, which this actor never tries to get past. Open a `sogou_link` in your browser to read the article.

#### Stable fields for automations

These fields were present in every row of their type in the test runs (77 search results and 5 articles). Their names will not change.

| Field | Row type | What it holds |
|---|---|---|
| `type` | all | search\_result or article |
| `title` | all | Article title |
| `account_name` | all | Official Account name |
| `published_at` | all | Publish time, ISO 8601 UTC |
| `cover_image_url` | all | Cover image |
| `input` | all | The input line the row answers |
| `scraped_at` | all | When the row was saved |
| `keyword` | search\_result | The keyword searched |
| `position` | search\_result | Position across the keyword's pages |
| `snippet` | search\_result | Sogou's excerpt |
| `sogou_link` | search\_result | Sogou's link to the article |
| `url` | article | The article link read |
| `account_gh_id` | article | The account's gh\_ id |
| `content_length` | article | Characters in the full text |
| `published_at_china_time` | article | Publish time in China time |

#### Output examples

A search result for 新能源汽车 (new energy vehicles):

```json
{
  "type": "search_result",
  "keyword": "新能源汽车",
  "page": 1,
  "position": 3,
  "title": "2026 年新能源汽车购车补贴政策深度解读及申领实操指南",
  "snippet": "进入 2026 年,我国新能源汽车消费支持政策完成新一轮优化调整,从过去的定额直补转向 ＂比例补贴 + 分级封顶＂ 的精准化模式,...",
  "account_name": "清水一瓢",
  "published_at": "2026-10-01T07:55:41.000Z",
  "cover_image_url": "https://mmbiz.qpic.cn/sz_mmbiz_jpg/ibW9I1oake7aZGTU4fj4k32RJKM3dWAiaoADPj3ay8YgKq3dc40TrAictricQnibm49MSH27ia4ejrltias4nbWgrgjCqyCDsK9Jia1uZsjdNdX0LrU/0?wx_fmt=jpeg",
  "sogou_link": "https://weixin.sogou.com/link?url=dn9a_-gY295K0Rci_xozVXfdMkSQTLW6cwJThYulHEtVjX...",
  "sogou_doc_id": "ab735a258a90e8e1-6bee54fcbd896b2a-44b71c1d093bea16a771651bcdc4fd0a",
  "total_results_estimate": 20241
}
```

A search result for 人工智能 (artificial intelligence):

```json
{
  "type": "search_result",
  "keyword": "人工智能",
  "page": 1,
  "position": 1,
  "title": "10月首批!重庆AI智能体培训启动,现已开始,助力冲刺大厂高薪offer,20岁以上可学",
  "account_name": "重庆潮生活",
  "published_at": "2026-10-01T03:03:41.000Z",
  "total_results_estimate": 32247
}
```

An article with the Original badge, text trimmed here:

```json
{
  "type": "article",
  "url": "https://mp.weixin.qq.com/s/f9uC5ztnJvNy5jaxQL2xQg",
  "title": "OPC时代，哪些生意会被重写",
  "account_name": "侠说",
  "account_wechat_id": "aTalkMan",
  "account_gh_id": "gh_80727f2e7eca",
  "account_biz": "MzkwNDY5MzAzNQ==",
  "account_signature": "全行业报告库，实战营销干货。6200+会员，6.9万+报告，保持日更新~",
  "author": "郭太侠",
  "description": "Agent替你干活，协议替你直连，信用替你背书，个人第一次成为结构的原点。",
  "published_at": "2026-09-22T23:00:00.000Z",
  "published_at_china_time": "2026-09-23 07:00",
  "publisher_region": "广东",
  "publisher_country": "中国",
  "is_original": true,
  "content_text": "hi，我是太侠，行业智库《侠说》主理人，内含6.9万行业报告，点击上方图片可下载报告。\n本篇报告内容拆解如下：\n\n你有没有发现，我们的大部分生意，中间都站着一个中间人。...",
  "content_length": 3257,
  "content_truncated": false,
  "image_count": 7,
  "video_count": 0
}
```

A media account's article:

```json
{
  "type": "article",
  "url": "https://mp.weixin.qq.com/s/WDYIwOCaCOP-dgmV563HPQ",
  "title": "2026AI竞速：向上突破天花板，向下扎根产业链",
  "account_name": "央视财经",
  "account_wechat_id": "cctvyscj",
  "account_gh_id": "gh_9dc0e48d383a",
  "published_at_china_time": "2026-07-26 14:36",
  "publisher_region": "北京",
  "is_original": false,
  "content_length": 3445,
  "image_count": 8,
  "video_count": 6
}
```

### Pricing

Pay per event: you pay for rows delivered, plus a flat start fee.

| Event | Free | Bronze | Silver | Gold and above |
|---|---|---|---|---|
| Row delivered, search result or article (per 1,000) | $3.99 | $3.49 | $2.99 | $2.49 |
| Run start (once per run) | $0.005 | $0.005 | $0.005 | $0.005 |

A search result and a full article cost the same, plus a flat $0.005 per run whatever memory you choose.

Never charged:

- a result already delivered in the same run (it showed up under a second keyword or on two pages);
- results and articles delivered in an earlier run when `onlyNewArticles` is on;
- articles WeChat no longer shows (deleted, removed, account moved, link expired);
- pages Sogou or WeChat refused;
- the info rows (`_type: "info"`): the note that explains a run that delivered nothing, and the closing note at the end of a run that did;
- a run whose input has nothing to scrape (not even the start fee).

You can cap spending in Apify Console with the run's maximum charge; the actor stops cleanly when it is reached.

### FAQ

#### What are WeChat Official Accounts?

WeChat Official Accounts (微信公众号) are the publishing channels inside WeChat used by companies, media, government bodies and writers in China. Their articles are the main long form content in Chinese social media, and most of them can be opened by anyone through an mp.weixin.qq.com link.

#### How many results can I get per keyword?

Up to 100: Sogou shows 10 results a page and 10 pages to a visitor. `total_results_estimate` tells you how many Sogou counts (20,241 for 新能源汽车 in our test), so for wider coverage use more and narrower keywords.

#### Why do search results not include the mp.weixin.qq.com link?

Sogou hides the article link behind its own redirect (`sogou_link`), which only resolves for a person clicking in a browser. An automated request for it is sent to Sogou's verification page, and this actor never tries to get past verification pages. So search rows give you the title, account, time, excerpt and cover, and `sogou_link` opens the article in your browser. When you have article links from anywhere else, `articleUrls` returns the full text.

#### Can I get read counts, likes or comments?

No. Read counts, likes, 在看 and comments are shown only inside the WeChat app to logged in users, not on the public article page. Tools that offer them use logged in WeChat accounts or paid data resellers; this actor uses neither.

#### Do I need a WeChat account, cookies or a phone?

No. It reads Sogou's public search and public article pages without logging in and without any cookies from you.

#### How fresh is the data?

Live. Every run reads Sogou and WeChat at that moment. In our test the newest results were published the same morning.

#### What happens with deleted or moved articles?

WeChat shows a notice instead of the article. The actor reads the notice and reports it in the run summary (for example "the account has moved" or "deleted by the publisher"). No row is delivered and nothing is charged.

#### What does `publisher_region` mean?

Since 2022 WeChat prints the region an article was published from under its title (for example 北京 or 广东), taken from the publisher's IP address. The actor returns what WeChat shows, with the country in `publisher_country`.

#### What proxy does it use?

Apify's datacenter proxy, included in your plan; you do not need to set anything. In our tests all 18 requests from datacenter addresses were answered without a verification page. If Sogou or WeChat ever answers with one, the actor does not try to solve it: the page counts as refused, it tries later from a new address, and it stops the run if refusals keep coming, without charging for anything not delivered.

#### Can I run it on a schedule and get only new articles?

Yes. Create a schedule in Apify Console and turn on `onlyNewArticles`. The actor remembers up to 50,000 results and articles per input and delivers only new ones.

#### What formats can I export?

JSON, CSV, Excel, XML, HTML and RSS from the dataset, or straight into Google Sheets, a webhook or the API.

#### Can I use it from Claude, ChatGPT or another AI assistant?

- Connector URL: `https://mcp.apify.com/?tools=themineworks/wechat-article-scraper`.
- Claude: Settings > Connectors > Add custom connector, paste the URL, sign in with Apify.
- ChatGPT: developer mode, add an MCP connector with the URL, sign in with Apify.
- Cursor or VS Code: add it as an HTTP MCP server with that URL.
- Claude Code: `claude mcp add -t http wechat-article-scraper "https://mcp.apify.com/?tools=themineworks/wechat-article-scraper"`.

#### Is it legal to scrape WeChat articles?

The actor reads only public pages anyone can open without logging in: Sogou's WeChat search results and public article pages. It does not collect readers' data, comments or private messages. Note that the robots.txt files of weixin.sogou.com and mp.weixin.qq.com ask automated crawlers not to fetch these pages. Articles are copyrighted by their publishers. You are responsible for how you use the data, including the sites' terms, copyright, and data protection laws such as GDPR, CCPA and China's PIPL.

### Integrations

- **Google Sheets:** send each run's rows to a sheet with Apify's Google Sheets integration.
- **Make, Zapier and n8n:** start a run and pick up new articles in your own flow.
- **Webhooks:** get a call when a run finishes, with the dataset link.
- **API and SDKs:** run it from Python or JavaScript with the Apify client, as in the block above.
- **MCP clients:** Claude, Cursor and other MCP clients can call it through mcp.apify.com.

### More from The Mine Works

**More tools**

- [G2 Reviews Scraper](https://apify.com/themineworks/g2-reviews-scraper)
- [Tennis Match & Player Data Scraper](https://apify.com/themineworks/tennis-match-data)
- [Flashscore Tennis Scraper](https://apify.com/themineworks/flashscore-tennis-results-scraper)
- [LandWatch Scraper](https://apify.com/themineworks/landwatch-land-for-sale-scraper)
- [Capterra Reviews Scraper](https://apify.com/themineworks/capterra-software-reviews-scraper)
- [Google Hotels Prices Scraper](https://apify.com/themineworks/google-hotels-prices-scraper)
- [Taobao Products Scraper 淘宝 天猫](https://apify.com/themineworks/taobao-products-scraper)
- [Google Lens OCR Scraper](https://apify.com/themineworks/google-lens-ocr-scraper)

**Social media and video**

- [Threads Scraper](https://apify.com/themineworks/threads-scraper)
- [Reddit Scraper](https://apify.com/themineworks/reddit-scraper)
- [Threads Search Scraper](https://apify.com/themineworks/threads-search-scraper)
- [Instagram Profile Scraper](https://apify.com/themineworks/instagram-profile-scraper)

**Leads and business directories**

- [B2B Leads Finder](https://apify.com/themineworks/b2b-leads-finder)
- [Skip Trace Lookup](https://apify.com/themineworks/skip-trace-lookup)
- [Google Maps Email Scraper](https://apify.com/themineworks/maps-leads)
- [JustDial Scraper](https://apify.com/themineworks/justdial-business)

**Marketing, SEO and reviews**

- [Facebook Ad Library Scraper](https://apify.com/themineworks/meta-ad-library-scraper)
- [Google Ads Transparency Scraper](https://apify.com/themineworks/google-ads-transparency)
- [Similarweb Scraper](https://apify.com/themineworks/similarweb-scraper)
- [Google News Scraper](https://apify.com/themineworks/google-news)

**LinkedIn**

- [LinkedIn Company Scraper](https://apify.com/themineworks/linkedin-company-details)
- [LinkedIn Post Scraper](https://apify.com/themineworks/linkedin-post-search)
- [LinkedIn Employees Scraper](https://apify.com/themineworks/linkedin-employees)
- [LinkedIn Profile Scraper](https://apify.com/themineworks/linkedin-profile-scraper)

**Real estate**

- [Zillow Rentals Scraper](https://apify.com/themineworks/zillow-rental-listings)
- [Zillow Sold Comps Scraper](https://apify.com/themineworks/zillow-recently-sold)
- [Housing.com Scraper](https://apify.com/themineworks/housing-com-scraper)
- [India Real Estate MCP](https://apify.com/themineworks/india-real-estate-mcp)

**Science, health and government data**

- [CourtListener Scraper](https://apify.com/themineworks/courtlistener-court-records)
- [Socrata Open Data Scraper](https://apify.com/themineworks/socrata-open-data)
- [Academic Research MCP](https://apify.com/themineworks/academic-research-mcp)
- [OpenAlex Scraper](https://apify.com/themineworks/openalex-scholarly-works)

**Jobs and hiring**

- [Foundit Monster India Jobs](https://apify.com/themineworks/foundit-jobs-scraper)
- [Hirist Jobs Scraper](https://apify.com/themineworks/hirist-jobs-scraper)
- [India Jobs MCP](https://apify.com/themineworks/india-jobs-mcp)
- [Naukri Jobs Scraper](https://apify.com/themineworks/naukri-jobs)

**Company and business data**

- [GST Taxpayer Lookup](https://apify.com/themineworks/gst-taxpayer-lookup)
- [Company Domain Finder](https://apify.com/themineworks/company-domain-finder)
- [World Bank Trade Scraper](https://apify.com/themineworks/global-trade-data)
- [SEC EDGAR Filings Scraper](https://apify.com/themineworks/sec-edgar-filings)

**E-commerce and marketplaces**

- [Ozon.ru Scraper](https://apify.com/themineworks/ozon-product-search)
- [Amazon Product Scraper](https://apify.com/themineworks/amazon-products)
- [⭐ Amazon Reviews Scraper](https://apify.com/themineworks/amazon-reviews)
- [Carsales.com.au Scraper](https://apify.com/themineworks/carsales-scraper)

**Food and local services**

- [NoBroker Scraper](https://apify.com/themineworks/nobroker-scraper)
- [Swiggy Restaurant Scraper](https://apify.com/themineworks/swiggy-scraper)
- [Zomato Scraper](https://apify.com/themineworks/zomato-scraper)

**Developer and AI tools**

- [Website to Markdown Crawler](https://apify.com/themineworks/rag-crawler)
- [GitHub Repo Scraper](https://apify.com/themineworks/github-repo-intelligence)
- [GitHub Skill Finder](https://apify.com/themineworks/github-skill-discovery)
- [GitHub Trending Scraper](https://apify.com/themineworks/github-trending-scraper)

### Support

Found a problem or want a field added? Open an issue on the Issues tab and include the run link. For a new source, email dmineworks@gmail.com.

*WeChat Article Scraper finds WeChat Official Account articles by keyword through Sogou and returns the full text, account and publish details of any public article link, without a WeChat login.*

# Actor input Schema

## `keywords` (type: `array`):

One per line: words to search WeChat Official Account articles for (人工智能, 新能源汽车, a brand or person). Searched through Sogou's WeChat search, 10 results a page, at most 10 pages per keyword. A weixin.sogou.com search link works too. Up to 100 per run.

## `articleUrls` (type: `array`):

One per line: public WeChat article links, https://mp.weixin.qq.com/s/... or https://mp.weixin.qq.com/s?\_*biz=...\&mid=...\&idx=...\&sn=... Each gives one row with the full text, account name, WeChat id, gh* id, publish time and region, cover and image count. Up to 500 per run.

## `maxResultsPerKeyword` (type: `integer`):

Most search results delivered per keyword. Sogou shows 10 a page and at most 10 pages, so 100 is the ceiling.

## `includeFullText` (type: `boolean`):

Article rows carry the full text in content\_text. Off: everything but the text (smaller rows); the price per row is the same.

## `maxTextLength` (type: `integer`):

Article text longer than this is cut, and content\_truncated is true. Long reports run past 20,000 characters.

## `maxItems` (type: `integer`):

The run stops once this many rows are delivered.

## `onlyNewArticles` (type: `boolean`):

For schedules: remember every search result and article this input delivered and deliver only new ones next time. Rows seen before are skipped and never charged.

## Actor input object example

```json
{
  "keywords": [
    "人工智能"
  ],
  "maxResultsPerKeyword": 30,
  "includeFullText": true,
  "maxTextLength": 50000,
  "maxItems": 1000,
  "onlyNewArticles": false
}
```

# Actor output Schema

## `results` (type: `string`):

No description

## `summary` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "keywords": [
        "人工智能"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("themineworks/wechat-article-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "keywords": ["人工智能"] }

# Run the Actor and wait for it to finish
run = client.actor("themineworks/wechat-article-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "keywords": [
    "人工智能"
  ]
}' |
apify call themineworks/wechat-article-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,themineworks/wechat-article-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/6XA30YmC0puKLClUF/builds/Ohf0MxX0DWTtujNBU/openapi.json
