# Chinese AI Crawlers: Website Access Check

**Use case:** 

Check robots.txt and llms.txt signals for Chinese AI crawlers on two example domains. Replace the sites with your own domains to repeat the check.

## Input

```json
{
  "sites": [
    "https://example.com",
    "https://apify.com"
  ],
  "paths": [
    "/"
  ],
  "bots": [
    "Baiduspider",
    "Baidubot",
    "Bytespider",
    "Bytedance",
    "Doubao",
    "DeepSeekBot",
    "ChatGLM-Spider",
    "Qwenbot",
    "PanguBot",
    "PetalBot",
    "Kimi",
    "Hunyuan",
    "YiBot",
    "SenseBot",
    "iFlytekBot",
    "MiniMaxBot",
    "InternLMBot",
    "360Spider",
    "Sogou web spider",
    "Yisouspider"
  ],
  "checkLlms": true,
  "maxItems": 2
}
```

## Output

```json
{
  "website": {
    "label": "website",
    "format": "text"
  },
  "aiAccessScore": {
    "label": "aiAccessScore",
    "format": "text"
  },
  "aiAccessNumerator": {
    "label": "aiAccessNumerator",
    "format": "text"
  },
  "aiAccessDenominator": {
    "label": "aiAccessDenominator",
    "format": "text"
  },
  "aiAccessUnknown": {
    "label": "aiAccessUnknown",
    "format": "text"
  },
  "aiSearchScore": {
    "label": "aiSearchScore",
    "format": "text"
  },
  "aiSearchNumerator": {
    "label": "aiSearchNumerator",
    "format": "text"
  },
  "aiSearchDenominator": {
    "label": "aiSearchDenominator",
    "format": "text"
  },
  "aiSearchUnknown": {
    "label": "aiSearchUnknown",
    "format": "text"
  },
  "robotsTxt": {
    "label": "robotsTxt",
    "format": "text"
  },
  "llmsTxt": {
    "label": "llmsTxt",
    "format": "text"
  },
  "blockedBots": {
    "label": "blockedBots",
    "format": "text"
  },
  "summary": {
    "label": "summary",
    "format": "text"
  },
  "matrix": {
    "label": "matrix",
    "format": "text"
  },
  "recommendations": {
    "label": "recommendations",
    "format": "text"
  },
  "contentSignals": {
    "label": "contentSignals",
    "format": "text"
  },
  "catalogVersion": {
    "label": "catalogVersion",
    "format": "text"
  },
  "catalog": {
    "label": "catalog",
    "format": "text"
  },
  "unprocessed": {
    "label": "unprocessed",
    "format": "text"
  },
  "schemaVersion": {
    "label": "schemaVersion",
    "format": "text"
  },
  "type": {
    "label": "type",
    "format": "text"
  },
  "sourceUrl": {
    "label": "sourceUrl",
    "format": "text"
  },
  "found": {
    "label": "found",
    "format": "text"
  },
  "status": {
    "label": "status",
    "format": "text"
  },
  "resultCount": {
    "label": "resultCount",
    "format": "text"
  },
  "partial": {
    "label": "partial",
    "format": "text"
  },
  "error": {
    "label": "error",
    "format": "text"
  },
  "warnings": {
    "label": "warnings",
    "format": "text"
  },
  "checkedAt": {
    "label": "checkedAt",
    "format": "text"
  },
  "evidence": {
    "label": "evidence",
    "format": "text"
  },
  "confidence": {
    "label": "confidence",
    "format": "text"
  },
  "action": {
    "label": "action",
    "format": "text"
  }
}
```

## About this Actor

This example demonstrates how to use [Chinese AI Crawler Access Checker for robots.txt](https://apify.com/zinin/chinese-ai-crawler-access-checker.md) with a specific input configuration. Visit the [Actor detail page](https://apify.com/zinin/chinese-ai-crawler-access-checker.md) to learn more, explore other use cases, and run it yourself.


## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
This Task's input is already configured above — use it as-is rather than inventing a new one.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For full API examples (JavaScript, Python, CLI, MCP, OpenAPI), see this Task's Actor page: https://apify.com/zinin/chinese-ai-crawler-access-checker.md

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).
