# Weibo Scraper - Hot Search, Trending Posts, Comments & Profiles (`pipiagent/weibo-scraper`) Actor

Scrape Weibo (微博), China's largest microblog: the hot search chart, trending posts, post details, full comment threads and user profiles. English field names. No login or cookie needed.

- **URL**: https://apify.com/pipiagent/weibo-scraper.md
- **Developed by:** [Frank0306](https://apify.com/pipiagent) (community)
- **Categories:** Social media, News, AI
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$5.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Weibo Scraper

Extract public data from [Weibo](https://weibo.com) (微博), China's largest microblogging platform with more than 500 million monthly users. Weibo is where news breaks and where brands, celebrities and the public talk to each other in China.

Use it to track what is trending in China, to measure the reaction to a post or a campaign, to collect Chinese-language comments for sentiment analysis, or to check the reach of an influencer.

- **Full comment threads.** Collect the comments under any public post, with replies, likes and the region each comment was posted from.
- **Hot search chart.** The live ranking of trending topics with heat scores and categories.
- **English field names.** Output is clean JSON with names like `likes`, `reposts` and `createdAt`. Text content stays in Chinese.
- **No login, no cookie and no proxy needed.** Press Start and get data.
- **Low price.** $5 per 1,000 results.

### What data can you get?

| Mode | One result is | Main fields |
|---|---|---|
| Hot search chart | One trending topic | Rank, topic, heat score, category, label, time on chart |
| Trending posts | One post | Text, likes, comments, reposts, images, video, author, region |
| Post details | One post | The same fields for the posts you list, with the full text of long posts |
| Comments | One comment or reply | Text, likes, reply count, date, region, author, author followers |
| User profiles | One account | Followers, following, post count, total likes received, verification, bio, location, company, school |

### How to use it

1. Click **Try for free**.
2. Choose a mode in **What to scrape**.
3. For post details and comments, paste post URLs. For profiles, paste profile URLs, user ids or account names.
4. Click **Start**, then download the results as JSON, CSV or Excel.

### Input examples

The hot search chart right now:

```json
{
    "mode": "trends"
}
```

Comments and replies of a post:

```json
{
    "mode": "comments",
    "postUrls": ["https://weibo.com/1669879400/RjCH5q8hs"],
    "includeReplies": true,
    "maxItems": 1000
}
```

Profiles of accounts:

```json
{
    "mode": "profiles",
    "users": ["人民日报", "https://weibo.com/u/1669879400"]
}
```

### Output examples

A trending topic:

```json
{
    "type": "trend",
    "chart": "hotSearch",
    "rank": 1,
    "topic": "房贷贴息1个百分点",
    "heat": 3020516,
    "category": "财经",
    "label": "新",
    "onChartSince": "2026-09-29T10:20:21Z",
    "url": "https://s.weibo.com/weibo?q=%23%E6%88%BF%E8%B4%B7%E8%B4%B4%E6%81%AF1%E4%B8%AA%E7%99%BE%E5%88%86%E7%82%B9%23",
    "scrapedAt": "2026-09-29T11:20:05Z"
}
```

A comment:

```json
{
    "type": "comment",
    "commentId": "5346770336090639",
    "postId": "5346769756228362",
    "postUrl": "https://weibo.com/1669879400/RjCH5q8hs",
    "isReply": false,
    "parentCommentId": null,
    "text": "迪迪抱抱[期待][期待]",
    "createdAt": "2026-09-24T12:27:29Z",
    "likes": 18810,
    "replyCount": 1474,
    "postedFrom": "江苏",
    "likedByPostAuthor": false,
    "authorId": "1787569845",
    "authorName": "护舒宝",
    "authorVerified": true,
    "authorVerifiedType": "company",
    "authorFollowers": 619498,
    "authorCity": "北京"
}
```

A user profile:

```json
{
    "type": "user",
    "userId": "2803301701",
    "name": "人民日报",
    "url": "https://weibo.com/u/2803301701",
    "bio": "人民日报法人微博。参与、沟通、记录时代。",
    "location": "北京",
    "followers": 158011657,
    "following": 3096,
    "posts": 153707,
    "totalLikesReceived": 2527373228,
    "isVerified": true,
    "verifiedType": "media",
    "verifiedReason": "《人民日报》法人微博",
    "joinedAt": "2012-07-22 02:28:35"
}
```

Dates are in UTC, except `joinedAt`, which is in China Standard Time (UTC+8). Emoticons appear as codes in square brackets, for example `[期待]`.

### How much does it cost?

You pay per saved result: **$5.00 per 1,000 results**. There is no start fee and no monthly rental.

| Run | Results | Cost |
|---|---|---|
| The hot search chart | about 50 | $0.25 |
| 1,000 comments of a post | 1,000 | $5.00 |
| 100 user profiles | 100 | $0.50 |

Set a maximum charge per run in the run options and the Actor stops when it is reached.

### Limits you should know

Weibo shows only part of its content to visitors who are not logged in. The Actor collects what a visitor can see:

- **Keyword search is not included.** Weibo requires login to search posts.
- **A user's post history is not included.** Weibo requires login to open the list of posts on a profile. You can still get any post when you have its URL.
- **Repost lists are not included.**
- **Comments are not always complete on very large threads.** Weibo stops serving more comments after some depth, which differs from post to post. The `comments` field of the post shows how many exist in total.
- **Trending posts are a sample.** The feed returns what Weibo recommends at that moment and changes on every run.

### Tips

- **Run the chart on a schedule.** Scrape the hot search chart every 10 or 30 minutes to build a history of what trended and for how long.
- **Combine modes.** Run trending posts first, then feed the `url` values into comments mode.

### FAQ

**Is it legal?** The Actor collects only data that Weibo shows publicly to every visitor. Results can contain personal data such as usernames. Make sure you have a legitimate reason to process it, as required by the GDPR and similar laws.

**Does it download images and videos?** No. It returns their URLs. Weibo media URLs can expire after some hours.

**Something is broken or missing?** Open a ticket in the Issues tab. Issues are usually answered within a day.

# Actor input Schema

## `mode` (type: `string`):

Choose the kind of data you need. Each mode uses the matching input below.

## `chart` (type: `string`):

Used in hot search chart mode.

## `postUrls` (type: `array`):

Used in post details and comments modes. Accepts weibo.com and m.weibo.cn post URLs, for example https://weibo.com/1669879400/RjCH5q8hs.

## `users` (type: `array`):

Used in user profiles mode. Accepts profile URLs, numeric user ids or account names, for example "人民日报".

## `maxItems` (type: `integer`):

Maximum number of results to save in this run. You are only charged for saved results.

## `maxItemsPerPost` (type: `integer`):

Used in comments mode. Optional limit for each post, so that one post with many comments does not use up the whole run.

## `commentSort` (type: `string`):

Used in comments mode.

## `includeReplies` (type: `boolean`):

Used in comments mode. Also saves the replies under each comment. Every reply counts as one result.

## `proxyConfiguration` (type: `object`):

Optional. The Actor works without a proxy. Turn one on only if runs start failing with blocked requests.

## Actor input object example

```json
{
  "mode": "trends",
  "chart": "hotSearch",
  "maxItems": 100,
  "commentSort": "hot",
  "includeReplies": false,
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# Actor output Schema

## `results` (type: `string`):

Trends, posts, comments or user profiles saved by the run, one item per result.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {};

// Run the Actor and wait for it to finish
const run = await client.actor("pipiagent/weibo-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {}

# Run the Actor and wait for it to finish
run = client.actor("pipiagent/weibo-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{}' |
apify call pipiagent/weibo-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,pipiagent/weibo-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/rWi11e4RR51gxxynX/builds/4kVBYfLKPYFXuvrKS/openapi.json
