Reddit Posts & Comments Scraper avatar

Reddit Posts & Comments Scraper

Pricing

Pay per event

Go to Apify Store
Reddit Posts & Comments Scraper

Reddit Posts & Comments Scraper

Extract Reddit posts and comments from any subreddit, search query, or user profile. Collect titles, scores, comments, media URLs, and 40+ fields per-post. Supports multiple subreddits, advanced filtering by score, flair, domain, and post type, plus optional comment enrichment.

Pricing

Pay per event

Rating

0.0

(0)

Developer

ParseForge

ParseForge

Maintained by Community

Actor stats

2

Bookmarked

340

Total users

37

Monthly active users

5 days ago

Last modified

Share

ParseForge Banner

๐Ÿ“ฑ Reddit Posts and Comments Scraper

๐Ÿš€ Pull Reddit posts and full comment trees in minutes. Subreddits, search, user profiles, multi-subreddits, post comments. Runs with no login by default; add optional Reddit credentials to unlock engagement metrics and comment trees.

Pull live Reddit posts and comments from any subreddit, search query, user profile, multi-subreddit feed, or specific post. The actor walks Reddit's public API surface for the mode you pick and returns one structured record per post or comment ready for trend analysis, content research, brand listening, or community studies.

Every run fetches data live so you get the current state of Reddit at run time, not a stale dump. Records include title, full text, author, subreddit, timestamps, and a back-reference URL by default, plus score, upvote ratio, comment count, awards, flair, and the full nested comment tree when you enable the optional rich-data mode.

๐Ÿ”“ Two read modes: no-login (default) and rich-data (optional)

Reddit no longer serves its unauthenticated JSON API to most automated traffic, so this Actor reads Reddit two ways and picks the right one for you automatically:

๐Ÿ“ฐ Default (no login)๐Ÿ”‘ Rich data (optional)
SetupNothing. Just run it.Add REDDIT_CLIENT_ID + REDDIT_CLIENT_SECRET
SourcePublic RSS/Atom feedsOfficial Reddit OAuth API
FieldsCore: title, text, author, subreddit, url, timestampsFull 40+ fields
Engagement (score, comment count, upvote ratio, awards)Not availableโœ… Included
Comment treesNot availableโœ… Included
Best forContent research, brand/keyword listening, trend discoveryAnalytics, ranking, sentiment, full datasets

No credentials? No problem. The Actor reads Reddit out of the box with zero setup. Engagement-only fields come back as 0/null in this mode.

Want the rich data? Create a free Reddit app (type script) at reddit.com/prefs/apps and add the client ID and secret as Actor environment variables. The Actor detects them and switches to the official API automatically, unlocking score, comment counts, and full comment trees.

๐Ÿ‘ฅ Built for๐ŸŽฏ Primary use cases
Brand and social listeningTrack Reddit mentions of your brand
Content research teamsMine Reddit for content ideas and angles
Market researchersStudy sentiment and discussion topics
Crisis monitoringWatch for reputation events in real time
Marketing and growthIdentify trending topics for content marketing
ResearchersStudy community dynamics and discussion patterns

๐Ÿ“‹ What the Reddit Posts and Comments Scraper does

  • ๐ŸŽฏ Five scraping modes. Subreddit, search, user profile, multi-subreddit, or specific post comments.
  • ๐Ÿ”“ No login by default. Reads public RSS feeds with zero setup; add optional Reddit credentials for the full dataset.
  • ๐Ÿ“Š Core metadata always. Title, body, author, subreddit, url, timestamps.
  • ๐Ÿ“ˆ Engagement with credentials. Score, upvote ratio, comment count, awards, and flair when you enable the optional rich-data mode.
  • ๐Ÿ’ฌ Comment trees (rich-data mode). Nested comment tree (with includeComments) or comments-only mode.
  • ๐Ÿ” Filter by metadata. Min score, post type (text / link / image / video), flair, domain (rich-data mode).
  • ๐Ÿ—“๏ธ Time filters. All time, year, month, week, day, hour.

The scraper accepts a mode plus the matching inputs (subreddit name, search query, username, list of subreddits, or post URL). It walks Reddit's public surface and pushes structured records to the dataset as posts are processed.

๐Ÿ’ก Why it matters: Reddit is the largest public discussion forum on the web but its UI lacks bulk export. A live, structured pull beats manual scraping for brand listening, content research, and trend analysis.

๐Ÿ“Š Data fields

Each record includes: author, authorFullname, contentCategories, createdAt, createdUtc, domain, editedAt, fullname, galleryImages, gilded, id, isNsfw, isOriginalContent, isSelf, isSpoiler, isStickied, isVideo, linkFlairCssClass, linkFlairText, mediaUrl, numComments, numCrossposts, permalink, postHint, previewHeight, previewImage, previewWidth, score, scrapedAt, selftext, selftextHtml, subreddit, subredditPrefixed, suggestedSort, thumbnail, thumbnailHeight, thumbnailWidth, title, totalAwards, upvoteRatio. All 40 field names come from a real production run, so what you see here is what lands in your dataset.

โš ๏ธ Good to Know: in comments mode, the dataset returns one record per comment (not per post). Switch to subreddit mode with includeComments: true if you want both posts and their top comments in the same run.

๐Ÿš€ How to use

  1. ๐Ÿ“ Sign up. Create a free account with $5 credit (takes 2 minutes).
  2. ๐ŸŒ Open the Actor. Go to the Reddit Posts and Comments Scraper page on the Apify Store.
  3. ๐ŸŽฏ Pick mode. Choose subreddit, search, user, multi, or comments mode and set the matching inputs.
  4. ๐Ÿš€ Run it. Click Start and let the Actor collect your data.
  5. ๐Ÿ“ฅ Download. Grab your results in the Dataset tab as CSV, Excel, JSON, or XML.

โฑ๏ธ Total time from signup to downloaded dataset: 3-5 minutes. No coding required.

๐Ÿ’ก Pro Tip: browse the complete ParseForge collection for more reference-data scrapers.

โš ๏ธ Disclaimer. This Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by Reddit. All trademarks mentioned are the property of their respective owners. The scraper accesses only publicly available pages and is intended for legitimate research, analytics, and brand-listening use. Users are responsible for compliance with Reddit's API Terms of Use, applicable privacy laws, and any data-protection rules that apply.

๐Ÿ†˜ Need Help?

If you hit a bug, have questions about setup, or need a scraper we haven't built yet, open our contact form or write to parseforge@protonmail.com. We also take on paid custom data projects.

For faster answers, join our Discord. It's the best place to get support and suggest new actors.