Subreddit Scraper - Reddit Posts Past 1,000, Dates, No Login avatar

Subreddit Scraper - Reddit Posts Past 1,000, Dates, No Login

Pricing

from $2.55 / 1,000 posts

Go to Apify Store
Subreddit Scraper - Reddit Posts Past 1,000, Dates, No Login

Subreddit Scraper - Reddit Posts Past 1,000, Dates, No Login

Returns the posts of any list of subreddits: title, text, score, comment count, author, flair and media links. Reads past the roughly 950 posts one Reddit list shows by combining its orders, cuts out a date range, and has an only-new mode for monitoring. No login, no API key.

Pricing

from $2.55 / 1,000 posts

Rating

0.0

(0)

Developer

Ben

Ben

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

4 hours ago

Last modified

Share

๐Ÿ“‹ Subreddit Scraper

Returns the posts of any list of subreddits without a Reddit login or API key. One row per post: title, text, score, upvote ratio, comment count, author, flair, media links and date. It reads past the roughly 950 posts one Reddit list shows, cuts out a date range, and has an only-new mode that turns a schedule into a feed of fresh posts.

Price: until October 18, 2026 $3.00 per 1,000 posts on the Apify Free plan ($2.55 from Gold up); from October 19 $1.50 ($1.20 from Gold up). A subreddit that is private or does not exist costs nothing. Export to JSON, CSV or Excel, run on a schedule, call via API, or connect to Make, Zapier or n8n.

๐Ÿ”Ž What is the Subreddit Scraper?

It reads subreddit feeds through Reddit's API, so no browser runs and 512 MB of memory is enough. A default run of ten posts takes about six seconds; a thousand posts take about a minute.

Reddit ends every feed early. The "new" feed of r/python stopped after 950 posts, which reached back seven and a half months. The same subreddit's other feeds (top of all time, of the year, of the month, hot, controversial, rising) hold partly other posts, so when you ask for more than one feed gives, the Actor reads them as well and saves every post once: 2,452 different posts for r/python in a test from our own server.

What data does it extract?

  • Post: id, title, text (plain and Markdown), link and permalink, flair, type flags (self, video, adult, spoiler, pinned, locked)
  • Engagement: score, upvote ratio, number of comments, awards
  • Author and subreddit
  • Media: preview images with sizes, thumbnail, linked domain
  • Dates: creation time (UTC)
  • For AI use: word count and estimated token count of the text

What users of Reddit scrapers ask for, and what this Actor does

The issue pages of the four most used Reddit Actors on Apify were read on October 4, 2026:

Asked for thereHere
"Max post limit is 900", "what are the actual limits per subreddit?"Reddit's other feeds are read as well: 2,452 different posts of one subreddit instead of 950
A start and end datepostedAfter and postedBefore. The newest-first feed is read first and left when the date is reached: August of two subreddits came back as 136 posts in 19 seconds
Several subreddits in one runsubreddits is a list; each gets its own limit
Runs blocked by Reddit (403), slow runsReads Reddit's API instead of its web pages; a private or missing subreddit is named in the status message and the run still succeeds
Bills that grow beyond the limit, double chargesA limit is exact, a post is saved once, and a run stops at its maximum charge with what it has saved

โฌ‡๏ธ Input

FieldTypeDefaultWhat it does
subredditsarrayNames or links: python, r/python, https://www.reddit.com/r/python/
sortstringhothot, new, top, rising, controversial: the first feed that is read
timeFilterstringallFor top and controversial: hour, day, week, month, year, all
maxPostsPerSubredditinteger25Posts per subreddit, up to 10,000
postedAfterstringA day (YYYY-MM-DD) or a period back from now (12 hours, 7 days)
postedBeforestringUpper end of a date range (YYYY-MM-DD)
onlyNewbooleanfalseSkip everything an earlier run with the same monitor name delivered
monitorIdstringName of that memory

Example input

This week's top posts of three subreddits:

{
"subreddits": ["python", "webscraping", "dataengineering"],
"sort": "top",
"timeFilter": "week",
"maxPostsPerSubreddit": 50
}

Everything posted in August:

{
"subreddits": ["webscraping"],
"postedAfter": "2026-08-01",
"postedBefore": "2026-08-31",
"maxPostsPerSubreddit": 2000
}

As many posts of a subreddit as Reddit gives:

{
"subreddits": ["python"],
"sort": "new",
"maxPostsPerSubreddit": 10000
}

A feed of new posts for a schedule:

{
"subreddits": ["AskReddit", "webscraping"],
"maxPostsPerSubreddit": 100,
"onlyNew": true,
"monitorId": "my-subreddits"
}

Input names of other Reddit Actors are read as well: subredditUrls, startUrls (subreddit links), subreddit, maxItems, maxPostCount, maxPostsCount, postDateLimit and sinceDate.

โฌ†๏ธ Output

One row per post (the text is shortened here):

{
"id": "1wx5qzu",
"title": "I built a site that logs and classifies scrapers that visit it.",
"url": "https://www.reddit.com/r/webscraping/comments/1wx5qzu/i_built_a_site_that_logs_and_classifies_scrapers/",
"permalink": "https://www.reddit.com/r/webscraping/comments/1wx5qzu/i_built_a_site_that_logs_and_classifies_scrapers/",
"selftext": "I run The Crawler Zoo, a site that identifies every bot that visits and puts it on display as a live exhibit. Since this sub builds the things it โ€ฆ",
"selftext_markdown": "I run The Crawler Zoo, a site that identifies every bot that visits and puts it on display as a live exhibit. Since this sub builds the things it โ€ฆ",
"author": "Time_Instruction_955",
"subreddit": "webscraping",
"subreddit_id": "t5_318ly",
"score": 12,
"upvote_ratio": 1,
"num_comments": 1,
"is_self": true,
"is_video": false,
"post_hint": null,
"domain": "self.webscraping",
"thumbnail": null,
"images": [],
"created_utc": "2026-10-04T03:38:43",
"total_awards_received": 0,
"link_flair_text": "Bot detection ๐Ÿค–",
"over_18": false,
"spoiler": false,
"stickied": false,
"locked": false,
"word_count": 294,
"token_count": 447,
"scraped_at": "2026-10-04T15:08:18.526979+00:00"
}

The run also writes a SUMMARY record with the number of posts per subreddit and the subreddits that were not available.

๐Ÿ’ฐ What a run costs

Apify planPer 1,000 posts until October 18, 2026From October 19, 2026
Free$3.00$1.50
Bronze$2.85$1.40
Silver$2.70$1.30
Gold, Platinum, Diamond$2.55$1.20

Apify's standard start event ($0.00005) is the only other charge; proxies are included. Not charged: a subreddit that is private or does not exist, a post outside your date limits, and a post already delivered in only-new mode.

โฑ๏ธ Measured on Apify (October 4, 2026)

RunPostsTime
Sample input (one subreddit)106 s
The first 958 posts of one subreddit95858 s
August of two subreddits13619 s
Only-new watch on two subreddits, second run right after the first1 new8 s

All at 512 MB.

๐Ÿ”” A feed of new posts

Turn on onlyNew, give the watch a monitorId and put the run on a schedule. The first run delivers the newest posts up to your limit. Later runs read each subreddit newest first and leave it at the first page without anything new, so a quiet subreddit costs nothing. Posts from before the oldest post of the first run stay outside the watch. The memory is a key-value store named reddit-subreddit-monitor in your own account; delete a record there to start over.

๐Ÿค– For AI agents

Smallest useful call:

{"subreddits": ["python"], "sort": "top", "timeFilter": "week", "maxPostsPerSubreddit": 25}

Each row is one post with title, selftext, score, num_comments, author, subreddit, created_utc and url. subreddits takes several entries. Add "postedAfter": "7 days" for recent posts. A subreddit that is not available returns no rows; the status message and the SUMMARY record name it. No credentials are needed.

๐Ÿ’ก Use cases

  • ๐Ÿ“ก Community monitoring: every new post of the subreddits your customers use.
  • ๐Ÿง  Datasets for models: thousands of posts with text, scores and token counts.
  • ๐Ÿ“Š Trend research: what a community voted to the top this week, this month, this year.
  • ๐Ÿ—‚๏ธ Archives of a period: all posts between two dates.

โš ๏ธ Limits, stated plainly

  • Reddit's feeds are the source. A large subreddit gives about 2,000 to 2,500 different posts over all of its feeds; a full history of every post is not available this way. For that, see the Reddit Archive Scraper.
  • Beyond the first feed the order is mixed. The first feed keeps the order you chose; posts from the further feeds follow it.
  • A date range reaches as far as the "new" feed. In a busy subreddit that feed covers days, in a quiet one months. Older ranges are filled from the top and controversial feeds only as far as those posts appear there.
  • Posts only. For the comments, pass a row's url to the Reddit Comments Scraper.
  • Scores are the values at the time of the run. Deleted and removed posts are not returned.

โ“ FAQ

Do I need a Reddit account or API key? No.

How many posts can I get from one subreddit? About 950 from one feed, and about 2,000 to 2,500 when the Actor combines the feeds. Set maxPostsPerSubreddit to the number you want.

How do I get only recent posts? Set postedAfter to a day or a period such as 7 days.

Can I read private subreddits? No. They are reported as not available and cost nothing.

Can I run it every hour? Yes, with onlyNew. A subreddit without new posts is left after one request.

How do I call it from code? With the Apify API or the Python and JavaScript clients; every run returns a dataset you can fetch as JSON or CSV. It also works as a tool through Apify's MCP server.

Is it legal? The Actor reads public Reddit posts. Posts carry usernames and can contain personal data: GDPR, CCPA and similar rules apply to how you store and use them, and Reddit's terms apply to you as well.

๐Ÿ”— You might also like

Keywords: subreddit scraper, reddit posts scraper, scrape subreddit, reddit subreddit posts, reddit feed export, reddit posts by date, reddit top posts, subreddit monitoring, reddit data for ai, reddit posts csv, no api key reddit scraper, reddit community data