# Reddit Comments Scraper - Full Threads, Replies & Depth (`headply/reddit-comments-scraper`) Actor

Export full Reddit comment threads from post URLs: every reply with depth, parent id, score and time, "load more" expanded. Or get comments for a whole subreddit or search. No login, no API key.

- **URL**: https://apify.com/headply/reddit-comments-scraper.md
- **Developed by:** [Mayowa Ogedengbe](https://apify.com/headply) (community)
- **Categories:** Social media, AI
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.95 / 1,000 post scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

An Actor is a serverless cloud program that runs on the Apify platform. It has two run modes.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.

Apify vocabulary and the platform model are defined once, in the agent quickstart at https://apify.com/agents.md.

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.

Do not guess an integration path. Every one of them is in the agent quickstart at https://apify.com/agents.md: the Apify MCP server, Agent Skills with the Apify CLI, the JavaScript and Python clients, the REST API, and the account-free path for an agent with no human to sign in. It also carries the rule on stating cost before the first paid run.

For examples already wired to this Actor's own input schema, see the [API](#api) section below.

Each client library has reference documentation the quickstart does not restate: [JavaScript/TypeScript](https://docs.apify.com/api/client/js/docs.md) (`npm install apify-client`) and [Python](https://docs.apify.com/api/client/python/docs.md) (`pip install apify-client`).

# README

## Reddit Comments Scraper - Full Threads, Replies & Depth

Get **every comment and reply** from any Reddit post, with who replied to whom, ready for a spreadsheet, a
research project or an AI pipeline. Hidden "load more comments" and "more replies" are opened for you, so you get
the whole conversation, not just the first screen.

### Why people use it

- ✅ **No Reddit account, no login, no API key.**
- 🌳 **The full conversation tree.** Each comment shows its depth and the comment it answers, so you can rebuild the
  thread or keep only top-level replies.
- 📦 **Many posts at once.** Paste a list of post links, or point it at a whole subreddit or search and get the
  comments of every post.
- 🔁 **Follow a thread over time.** On a schedule it returns only the comments added since the last run.
- 💸 **From $1.40 per 1,000 comments.**

### Quick start

1. Paste one or more post links into **Post URLs**:
   ```
   https://www.reddit.com/r/AskReddit/comments/1wwnme0/
   https://redd.it/1wwy2o0
   ```
2. Choose **Max comments per post** (default 200) and the **Comment sort**: Top, New, Best, Controversial, Old or Q\&A.
3. Press **Start**. You get the post first, then its comments.

**Comments for a whole subreddit or search:** add `r/subredditname` or a keyword in the optional section, keep
**Include comments** on, and choose how many posts to cover.

### What you get for each comment

| Field | Example |
|---|---|
| Comment text | "The PSP ran a custom Sony OS on a MIPS CPU, so…" |
| Author | `Designer_Shelter7345` |
| Score | `3` |
| Posted at | `2026-10-03T21:27:41Z` |
| Depth | `0` = reply to the post, `1` = reply to a comment, … |
| Replying to | The id of the comment (or post) it answers |
| Post id, subreddit | Which post and community it belongs to |
| Link | Direct link to the comment on Reddit |

Each post comes with its title, text, author, score, upvote ratio, comment count, flair and media links.

View the results in the **Posts** and **Comments** tables, or download them as JSON, CSV, Excel, XML or HTML.

### Pricing

Prices drop on higher Apify plans.

| What | Price per 1,000 (Free plan → Gold and above) |
|---|---|
| Comment | **$1.40** → $0.95 |
| Post | $1.40 → $0.95 |
| New comment since your last run (when following a thread) | $0.50 → $0.35 |
| Each run | $0.001 |

What that looks like: a 500-comment thread ≈ **$0.70**. 100 posts with 50 comments each (5,100 results) ≈ **$7.14**.

### Follow a thread

Switch on **Only new since last run**, set **Comment sort** to *New*, and click **Schedule**. Each run brings only
the comments posted since the previous one. This works well for AMAs, product launches, support threads and
anywhere your brand is discussed.

### Connect it to your tools

- **n8n:** add the **Apify** node → *Run an Actor and get dataset*, choose this Actor, and send the comments on to
  Google Sheets, a database, Slack or an AI step.
- **Make / Zapier:** use the Apify app's *Run Actor* action and map the results.
- **API:** see the *API* tab on this page for ready-made examples in several languages.

### Good to know

- **Very large threads** (thousands of comments) take longer. Set *Max comments per post* to what you need; with
  *Top* sort you get the most upvoted comments first.
- **Deleted or removed comments** appear with an empty author or text.
- **Private, banned or quarantined subreddits** can't be read. They're listed in the run summary.
- Only public information that anyone can see on Reddit is collected.

### More Reddit tools

- [Reddit RSS Feed Replacement](https://apify.com/headply/reddit-rss-feed): new posts from subreddits, users and
  searches on a schedule, or as an RSS feed link. Reddit's RSS ends 13 Nov 2026.
- [Reddit Search Scraper](https://apify.com/headply/reddit-search-scraper): search Reddit by keyword, or get alerts.

Questions or a thread that doesn't come through? Open the **Issues** tab and we'll take a look.

# Actor input Schema

## `postUrls` (type: `array`):

Reddit post links (https://www.reddit.com/r/<sub>/comments/<id>/<slug>/, redd.it/<id> or t3\_<id>). Returns the post and its comment tree, with depth, parent id and "load more" expansion up to the comment limit.

## `maxCommentsPerPost` (type: `integer`):

Upper bound on comments per post, across all depths. "Load more" and "more replies" are expanded until this is reached. 0 skips comments.

## `commentSort` (type: `string`):

Order in which comments are collected. With a comment limit, this decides which comments you get. Use "New" with feed mode to follow a thread.

## `startUrls` (type: `array`):

Reddit URLs or shorthand, one per line. Accepts subreddits (r/python, https://www.reddit.com/r/python/new/), users (u/spez), search URLs (https://www.reddit.com/search/?q=n8n\&sort=new) and post URLs. Old RSS links work too (https://www.reddit.com/r/python/new/.rss): paste the feed URLs your reader used. r/a+b multireddits are split into one feed per subreddit.

## `includeComments` (type: `boolean`):

Also scrape the comment tree of every post found in subreddit, user and search feeds (post URLs always include comments).

## `searchQueries` (type: `array`):

Keywords or phrases to search across Reddit posts. Each keyword is a separate search and every result is tagged with the keyword that found it. Reddit search operators such as "subreddit:python" or quotes work.

## `searchSubreddits` (type: `array`):

Optional. Run every keyword inside each of these subreddits (python or r/python) instead of across all of Reddit.

## `sort` (type: `string`):

Order for subreddit and user feeds when the URL does not say. "New" is what RSS feeds used.

## `time` (type: `string`):

Time window for "Top" feeds and for search. Ignored by the other sorts.

## `searchSort` (type: `string`):

Order of search results. "Newest first" suits monitoring a keyword on a schedule.

## `maxItemsPerTarget` (type: `integer`):

Upper bound on posts returned for each subreddit, user or search keyword. Pages are fetched until this is reached or the feed ends.

## `onlyNewSinceLastRun` (type: `boolean`):

For schedules. Returns and bills only items not delivered by an earlier run with the same monitor name, and keeps an RSS/Atom feed at a fixed link. The first run of a feed delivers the latest posts at the standard price; every later run returns only new items, at the lower new-item price.

## `monitorName` (type: `string`):

Label for this feed's memory and RSS file. Runs with the same name share memory; use one name per schedule or client to keep them separate. Change it to start fresh.

## `rssOutput` (type: `boolean`):

Also write the results as RSS 2.0 (FEED.xml) and Atom (FEED.atom) to the run's key-value store. In feed mode the feed is also kept at a fixed link (shown in the log and the run summary) that feed readers and n8n/Zapier RSS steps can follow.

## `maxConcurrency` (type: `integer`):

Parallel requests. The default is fast and gentle; raise it for large jobs.

## `maxRequestRetries` (type: `integer`):

A page that fails to load is retried automatically up to this many times before it is skipped and listed in the run summary.

## `proxyConfiguration` (type: `object`):

Leave as is. The default is the setting that works reliably with Reddit.

## Actor input object example

```json
{
  "postUrls": [
    "https://www.reddit.com/r/AskReddit/comments/1wwnme0/"
  ],
  "maxCommentsPerPost": 200,
  "commentSort": "top",
  "includeComments": true,
  "sort": "new",
  "time": "none",
  "searchSort": "new",
  "maxItemsPerTarget": 10,
  "onlyNewSinceLastRun": false,
  "monitorName": "default",
  "rssOutput": false,
  "maxConcurrency": 5,
  "maxRequestRetries": 5,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# Actor output Schema

## `dataset` (type: `string`):

Dataset containing all posts and comments from this run.

## `summary` (type: `string`):

Counts per target, skipped inputs, request success ratio, bandwidth and the stable RSS URL in feed mode.

## `rss` (type: `string`):

This run's items as RSS 2.0 (when RSS output is on).

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "postUrls": [
        "https://www.reddit.com/r/AskReddit/comments/1wwnme0/"
    ],
    "maxCommentsPerPost": 200,
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("headply/reddit-comments-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "postUrls": ["https://www.reddit.com/r/AskReddit/comments/1wwnme0/"],
    "maxCommentsPerPost": 200,
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("headply/reddit-comments-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "postUrls": [
    "https://www.reddit.com/r/AskReddit/comments/1wwnme0/"
  ],
  "maxCommentsPerPost": 200,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call headply/reddit-comments-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,headply/reddit-comments-scraper"
        }
    }
}
```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/ey0B5hhJYJJezBQIV/builds/BC74GiubGS2sPaCze/openapi.json
