Reddit Subreddit Posts Scraper avatar

Reddit Subreddit Posts Scraper

Pricing

from $1.29 / 1,000 posts

Go to Apify Store
Reddit Subreddit Posts Scraper

Reddit Subreddit Posts Scraper

Scrape posts from any subreddit feed — hot, new, top, rising, or controversial — with the full Reddit post object, plus optional comments. One subreddit, a list, or a TXT/CSV file.

Pricing

from $1.29 / 1,000 posts

Rating

0.0

(0)

Developer

ParseBird

ParseBird

Maintained by Community

Actor stats

1

Bookmarked

2

Total users

1

Monthly active users

3 days ago

Last modified

Categories

Share

Reddit Subreddit Posts Scraper

Reddit Subreddit Posts Scraper collects posts from any public subreddit feed — hot, new, top, rising, or controversial — and saves the full Reddit post object for each one, with optional comment threads.

Scrape up to 1,000 posts per subreddit from one community, a list, or a TXT/CSV file — about 97 raw Reddit fields per post (score, upvote ratio, comments, flair, media, awards, moderation flags) plus up to 1,000 comments per post with depth, parent ID, and reply counts.

ParseBird Reddit Suite   •  Posts, comments & subreddit data
Reddit Search Scraper
Posts & comments by keyword
Reddit Comments Search Scraper
Comments by keyword & subreddit
Reddit Subreddit Posts Scraper
You are here

Copy to your AI assistant

Copy this block into ChatGPT, Claude, Cursor, or any LLM to start using this actor.

Use the Apify actor "parsebird/reddit-subreddit-posts-scraper" to scrape posts from Reddit subreddit feeds. Call it with the apify-client package: client.actor("parsebird/reddit-subreddit-posts-scraper").call(run_input={"subreddit": "technology", "sort": "top", "timeFilter": "week", "maxPostsPerSubreddit": 100}), then read results from client.dataset(run["defaultDatasetId"]).iterate_items(). Inputs: subreddit (string, name or URL), subreddits (array of names or URLs), subredditsFile (string URL to a TXT or CSV file), sort ("hot" | "new" | "top" | "rising" | "controversial", default "hot"), timeFilter ("hour" | "day" | "week" | "month" | "year" | "all", default "all", used only by top and controversial), maxPostsPerSubreddit (integer 1-1000, default 100), includeComments (boolean, default false), maxCommentsPerPost (integer 1-1000, default 1000). At least one of subreddit, subreddits, subredditsFile is required. Post rows have _type "post", _subreddit, createdAt, editedAt, full_link and every field of Reddit's post object (title, author, score, ups, upvote_ratio, num_comments, created_utc, permalink, url, selftext, selftext_html, domain, is_self, over_18, spoiler, stickied, locked, link_flair_text, preview, media, gallery_data, media_metadata, post_hint, subreddit_subscribers and more). Comment rows have _type "comment", _post_id, _subreddit, id, author, score, depth, parentId, body, createdAt, editedAt, permalink, childCount, isDeleted, isRemoved, authorId, authorIcon. Pricing: pay per post and per comment. API docs: https://apify.com/parsebird/reddit-subreddit-posts-scraper/api — get an API token at https://console.apify.com/settings/integrations

What is Reddit Subreddit Posts Scraper?

Reddit Subreddit Posts Scraper is a subreddit scraper and Reddit API alternative that reads subreddit feeds exactly the way Reddit sorts them and saves every post as structured data. Enter a subreddit such as AskReddit, choose hot, new, top, rising, or controversial, and get each post's title, text, score, upvote ratio, comment count, flair, media, and author — plus the post's comments if you want them.

The easiest way to try it: open the Input tab, keep AskReddit, and click Start. A run with the example input finishes in under a minute.

What can Reddit Subreddit Posts Scraper do?

  • 📰 All five subreddit feeds — hot, new, top, rising, and controversial, with hour-to-all-time windows for top and controversial.
  • 📦 Full Reddit post object — every field Reddit's feed returns (about 97 on a text post, more on image, video, and gallery posts), minus a few viewer-only fields.
  • 💬 Optional comments — up to 1,000 comments per post in thread order, including hidden "load more" replies, with depth, parentId, and childCount so you can rebuild the threads.
  • 📋 Bulk input — one subreddit, a list, or an uploaded TXT/CSV file. Names, r/ names, and full URLs all work and can be mixed.
  • ⚡ Parallel scraping — several subreddits run at the same time.
  • 🛑 Clear errors — private, banned, quarantined, Premium-only, and non-existent subreddits are skipped with a message, and the rest of the run continues.

Because it runs on the Apify platform, you also get:

  • Scheduling — run it every hour or day with Apify Schedules to monitor communities.
  • API access — start runs and fetch results with the Apify API or the Python and JavaScript clients.
  • Integrations — send data to Google Sheets, Slack, Zapier, Make, n8n, or webhooks with Apify integrations.
  • Export formats — download results as JSON, CSV, Excel, XML, or HTML.

What data can you extract from a subreddit?

Post rows (_type: "post") — the full post object from Reddit's feed, plus a few helper fields:

GroupFields
Helpers_type, _subreddit, _status, createdAt, editedAt, full_link
Content and mediatitle, selftext, selftext_html, url, url_overridden_by_dest, thumbnail, preview, media, secure_media, media_embed, gallery_data, media_metadata, post_hint, is_video, is_gallery
Engagementscore, ups, downs, upvote_ratio, num_comments, num_crossposts, total_awards_received, gilded, all_awardings
Author and flairauthor, author_fullname, author_flair_text, author_premium, link_flair_text, link_flair_css_class, link_flair_richtext, link_flair_background_color
Subredditsubreddit, subreddit_id, subreddit_name_prefixed, subreddit_type, subreddit_subscribers
Moderation and statusstickied, pinned, locked, archived, removed_by_category, distinguished, contest_mode
IDs and linksid, name, permalink, domain, created_utc, edited
Flagsover_18, spoiler, is_self, is_original_content, is_meta, is_crosspostable, is_robot_indexable, hide_score, no_follow, send_replies

Viewer-only fields that describe the logged-out visitor rather than the post (saved, clicked, hidden, visited, likes, report fields) are left out.

Comment rows (_type: "comment", when Include comments is on):

FieldDescription
_post_id, _subredditThe post (t3_…) and subreddit the comment belongs to
id, author, body, scoreComment ID (t1_…), author (null if deleted), Markdown text, and score
depth, parentIdNesting level (0 = top-level) and parent comment ID (null for top-level)
childCountDirect replies Reddit returned or listed for the comment
createdAt, editedAt, permalinkDates (ISO 8601, UTC) and link
isStickied, isLocked, isScoreHidden, distinguishedAs, isArchivedModeration and status flags
isDeleted, isRemoved, isInitiallyCollapsed, collapsedReasonDeleted, removed, and collapsed state
isOPWhether the commenter wrote the post
authorId, authorAccountType, authorFlair, authorIcon, authorIsCakeDayAuthor details
controversiality, totalAwardsReceivedReddit's controversy flag and awards

How to scrape a subreddit

  1. Open Reddit Subreddit Posts Scraper and go to the Input tab.
  2. Enter a Subreddit (for example technology), add several under Subreddits, or upload a Subreddits file.
  3. Choose Sort by. For Top or Controversial, pick a Time filter.
  4. Set Max posts per subreddit.
  5. Optionally turn on Include comments and set Max comments per post.
  6. Click Start, then download the data from the Output tab as JSON, CSV, or Excel.

Input parameters

ParameterTypeRequiredDefaultDescription
subredditstringOne of the three—A subreddit name or URL
subredditsarrayOne of the three[]Several subreddit names or URLs
subredditsFilestringOne of the three—Uploaded or hosted TXT (one per line) or CSV (subreddit, sub or name column, otherwise the first column)
sortstringNohothot, new, top, rising, controversial
timeFilterstringNoallhour, day, week, month, year, all — only for top and controversial
maxPostsPerSubredditintegerNo100Posts per subreddit (1–1,000)
includeCommentsbooleanNofalseAlso save each post's comments
maxCommentsPerPostintegerNo1000Comments per post (1–1,000)
proxyConfigurationobjectNoResidential Apify ProxyNetwork settings. Keep the default

Sort orders: hot is Reddit's default mix of recency and activity; new is newest first; top is highest score in the time window; rising is posts gaining traction right now (a short feed, usually about 25 posts); controversial is posts with a close split of upvotes and downvotes.

Input examples

Top posts of the past week:

{ "subreddit": "technology", "sort": "top", "timeFilter": "week", "maxPostsPerSubreddit": 200 }

Rising posts to catch trends early:

{ "subreddit": "wallstreetbets", "sort": "rising" }

Posts with comments:

{
"subreddit": "AskReddit",
"sort": "top",
"timeFilter": "day",
"maxPostsPerSubreddit": 50,
"includeComments": true,
"maxCommentsPerPost": 200
}

Several subreddits in one run:

{
"subreddits": ["AskReddit", "r/technology", "https://www.reddit.com/r/science"],
"sort": "hot",
"maxPostsPerSubreddit": 100
}

Bulk from a file:

{ "subredditsFile": "https://example.com/my-subreddits.csv", "sort": "new", "maxPostsPerSubreddit": 500 }

Most controversial posts of all time:

{ "subreddit": "unpopularopinion", "sort": "controversial", "timeFilter": "all", "maxPostsPerSubreddit": 100 }

Output example

A post row (the most-used fields shown; the full row has about 97 fields):

{
"_type": "post",
"_subreddit": "AskReddit",
"_status": "found",
"title": "Mike Tyson once said \"social media made people too comfortable disrespecting others and not getting punched in the face for it\" What’s a real-life example of someone learning that lesson?",
"author": "Famism316",
"score": 18959,
"upvote_ratio": 0.94,
"num_comments": 2630,
"createdAt": "2026-09-27T12:29:58.000Z",
"editedAt": null,
"created_utc": 1790512198.0,
"permalink": "/r/AskReddit/comments/1wrj0s2/mike_tyson_once_said_social_media_made_people_too/",
"full_link": "https://www.reddit.com/r/AskReddit/comments/1wrj0s2/mike_tyson_once_said_social_media_made_people_too/",
"url": "https://www.reddit.com/r/AskReddit/comments/1wrj0s2/mike_tyson_once_said_social_media_made_people_too/",
"selftext": "",
"domain": "self.AskReddit",
"is_self": true,
"over_18": false,
"spoiler": false,
"stickied": false,
"locked": false,
"num_crossposts": 2,
"subreddit_subscribers": 59745441,
"link_flair_text": null,
"name": "t3_1wrj0s2",
"id": "1wrj0s2"
}

A comment row:

{
"_type": "comment",
"_post_id": "t3_1wrj0s2",
"_subreddit": "AskReddit",
"_status": "found",
"id": "t1_pcd2mid",
"author": "FuzzyMcBitty",
"score": 11922,
"depth": 0,
"parentId": null,
"body": "Buzz Aldrin punching a moon landing denier. \n\nhttps://www.history.com/this-day-in-history/september-9/buzz-aldrin-punches-moon-landing-conspiracy-theorist-bart-sibrel",
"createdAt": "2026-09-27T13:24:58.000Z",
"editedAt": null,
"permalink": "/r/AskReddit/comments/1wrj0s2/mike_tyson_once_said_social_media_made_people_too/pcd2mid/",
"isStickied": false,
"isLocked": false,
"isScoreHidden": false,
"distinguishedAs": null,
"authorFlair": null,
"isDeleted": false,
"childCount": 39,
"authorId": "t2_561ct",
"authorAccountType": "USER",
"authorIsCakeDay": false,
"authorIcon": "https://www.redditstatic.com/avatars/defaults/v2/avatar_default_5.png",
"isArchived": false,
"isRemoved": false,
"isOP": false,
"isInitiallyCollapsed": false,
"collapsedReason": null,
"controversiality": 0,
"totalAwardsReceived": 0
}

Use Reddit Subreddit Posts Scraper via API

Python (apify-client):

from apify_client import ApifyClient
client = ApifyClient("<YOUR_APIFY_TOKEN>")
run = client.actor("parsebird/reddit-subreddit-posts-scraper").call(run_input={
"subreddits": ["technology", "science"],
"sort": "top",
"timeFilter": "week",
"maxPostsPerSubreddit": 100,
"includeComments": True,
"maxCommentsPerPost": 50,
})
for row in client.dataset(run["defaultDatasetId"]).iterate_items():
if row["_type"] == "post":
print(row["_subreddit"], row["score"], row["title"][:80])

JavaScript (apify-client):

import { ApifyClient } from 'apify-client';
const client = new ApifyClient({ token: '<YOUR_APIFY_TOKEN>' });
const run = await client.actor('parsebird/reddit-subreddit-posts-scraper').call({
subreddits: ['technology', 'science'],
sort: 'top',
timeFilter: 'week',
maxPostsPerSubreddit: 100,
});
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.filter((r) => r._type === 'post').forEach((p) => console.log(p._subreddit, p.score, p.title));

Use cases for subreddit data

  • Community monitoring — schedule new or rising to catch fresh posts in the subreddits your brand or market lives in.
  • Trend and content research — pull the top posts of the week or month to see what a community cares about.
  • Sentiment and NLP datasets — collect posts and full comment threads for classification, summarization, or LLM evaluation.
  • Competitive research — compare engagement across competing communities.
  • Moderation and community analytics — track stickied, locked, removed, and controversial posts over time.

How it works

  1. The scraper reads each subreddit's public feed in pages of up to 100 posts, in the sort order and time window you chose.
  2. Several subreddits are processed in parallel. Each post is saved as it arrives.
  3. When Include comments is on, a second pass loads each post's comment thread (skipping posts with no comments), expands hidden replies, adds author details, and saves the comments in thread order.

How much does it cost to scrape Reddit subreddits?

Reddit Subreddit Posts Scraper uses pay-per-event pricing with two events. Comments are charged only when you turn them on.

EventApify planPrice per eventPrice per 1,000
post-scrapedFree$0.00189$1.89
post-scrapedBronze$0.00169$1.69
post-scrapedSilver$0.00149$1.49
post-scrapedGold$0.00129$1.29
comment-scrapedFree$0.00089$0.89
comment-scrapedBronze$0.00079$0.79
comment-scrapedSilver$0.00069$0.69
comment-scrapedGold$0.00059$0.59

For example, 100 posts with up to 50 comments each (5,000 comments) cost about $0.19 + $4.45 = $4.64 on the Free plan. The Apify Free plan includes monthly platform credits you can use to try it, and you can set a maximum spend per run so the scraper stops at your budget.

FAQ

How many posts can I get from one subreddit? Up to 1,000 per feed — that is where Reddit's own listings stop. rising is much shorter (usually about 25 posts), and top/controversial with a short window only return the posts from that window.

Does the time filter work with hot, new, or rising? No. Reddit only supports time windows on top and controversial, so the filter is ignored for the other feeds.

Which comments are included? Comments are loaded in Reddit's "top" order and saved in thread order (each comment followed by its replies), including replies hidden behind "load more comments", until maxCommentsPerPost is reached. Very deep "continue this thread" branches are not expanded.

Why was a subreddit skipped? Reddit refuses logged-out access to private, banned, quarantined, and Premium-only communities. The run log and status message say which ones were skipped and why; the other subreddits are still scraped.

What file formats work for bulk input? A .txt file with one subreddit per line, or a .csv file with a subreddit, sub, or name column (otherwise the first column is used). Upload it in the Console or paste a link to a hosted file.

Can I schedule it or use it from my code and AI tools? Yes. Save your input as a task, add a schedule, and connect integrations. Developers can use the API tab, and AI agents can call it through the Apify MCP server.

Something doesn't work or you need a feature? Open an issue in the Issues tab. We read and answer every report.

This scraper only collects publicly available posts and comments that anyone can see without logging in. It does not access private or restricted communities, private messages, or account data. Results can include personal data such as usernames, which regulations like the GDPR protect. Only collect personal data when you have a legitimate reason, and check Reddit's terms and your local laws if you are unsure. Read more in Apify's post Is web scraping legal?.

Other Reddit scrapers by ParseBird