# Changelog of Reddit Scraper (`automation-lab/reddit-scraper`) Actor

- **URL**: https://apify.com/automation-lab/reddit-scraper/changelog.md
- **Full Actor documentation**: https://apify.com/automation-lab/reddit-scraper.md

## Changelog

This changelog highlights releases that materially changed Reddit Scraper for users. Dates use UTC; small internal refactors and routine maintenance are intentionally omitted.

### 2026-10-05

*Reliability*

- Reduced memory pressure when processing large subreddit/search responses and multi-target batches, without reducing output fields or limits.
- Fixed valid RSS posts and comments mentioning rate limits being mistaken for blocked responses.

### 2026-10-04

*Reliability*

- Fatal batch errors now preserve failed-run status instead of passing through normal-success cleanup. Scraping output and charging remain unchanged.

### 2026-09-22

*Features*

- Added opt-in Standby HTTP endpoints for short subreddit and post/comment requests. Batch runs, dataset output, and existing per-event prices remain unchanged.

### 2026-09-21

*Reliability*

- Large multi-target and comment-heavy runs now retain only a bounded in-memory export sample while streaming the complete dataset, reducing the risk of memory-related failures without changing output.

### 2026-09-18

*Reliability*

- Runs with a custom short platform timeout now stop before the deadline and preserve already scraped posts instead of being terminated while processing later targets.

### 2026-08-31

*Fixes*

- Fixed AI-ready output so `jsonl-finetune` assistant messages and `rag-markdown` documents include available comment text and author metadata in the transformed post record. These formats no longer emit or charge separate comment records.

### 2026-08-17

*Fixes*

- Fixed `sort` and `timeFilter` inputs being ignored for plain subreddit URLs. Subreddit runs now honor the requested ranking and time period unless the URL explicitly specifies its own sort or time filter.

### 2026-07-13

*Reliability*

- Restored bounded best-effort post score, upvote-ratio, and comment-score enrichment when Reddit blocks direct public-page requests. The Actor now retries only the optional enrichment request through a US residential proxy while keeping RSS discovery direct and preserving valid RSS output if enrichment remains unavailable.
- Direct post URLs now also attempt vote enrichment from Reddit's public post/comment markup; fields remain best-effort when Reddit does not expose them.

### 2026-07-01

*Fixes*

- Improved runs when Reddit blocks or rate-limits its RSS pages. The Actor now reaches its alternative public-data fallback sooner instead of exhausting unnecessary retries.
- Reduced the time and recovery overhead for blocked subreddit and search targets while preserving the existing output format.

### 2026-06-16

*Reliability*

- Blocked RSS responses are no longer reported as successful runs with an unexplained empty dataset.
- Empty, private, unavailable, and invalid targets now produce a clear target-status dataset record when possible, so automations can distinguish "no accessible results" from an Actor failure.
- Improved non-RSS recovery for subreddit listings and Reddit searches.

### 2026-06-10

*Features*

- Post records now include the discovered Reddit comment count when comments are scraped.
- Added optional delivery of a bounded run summary to an Apify MCP connector. This is opt-in through the `mcpDestinationConnector` input and does not replace the normal dataset output.

### 2026-06-04

*Integrations*

- MCP clients and AI assistants can now pass a single Reddit URL through `url`, `comment_url`, or `commentUrl`; the Actor normalizes these aliases to the standard `urls` input.

### 2026-06-03

*Features*

- Added focused comment-context mode for a specific Reddit comment URL. It returns the target comment and the surrounding context Reddit makes publicly available.

### 2026-05-31

*Changes*

- Reworked Reddit recovery around public RSS and fallback pages, restoring practical coverage for listings, searches, users, posts, and comments after Reddit restricted older access paths.
- Removed internal recovery metadata from public dataset records. Some vote, award, subscriber, media, and deep-comment fields remain best-effort when Reddit does not expose them publicly.

### 2026-05-29

*Reliability*

- Expanded fallback coverage across subreddit listings, searches, user profiles, direct posts, and nested comments when Reddit limits its primary public pages.

### 2026-05-24

*Features*

- Added convenient subreddit and user shortcuts, finer comment controls, and AI-ready `jsonl-finetune` and `rag-markdown` output formats.

### 2026-05-14

*Safety*

- The Actor now respects Apify's maximum charge per run and stops cleanly when the configured spending cap is reached.

### 2026-05-11

*Fixes*

- Fixed request compatibility regressions that caused affected runs to finish with zero Reddit items. Users affected by empty results could rerun the same input after this release.

### 2026-04-05

*Features*

- Added `jsonl-finetune` output for supervised fine-tuning datasets and `rag-markdown` output for retrieval and LLM workflows.

### 2026-03-22

*Features*

- Added strict keyword filtering and duplicate removal for Reddit posts.

*Fixes*

- Improved handling of nullable Reddit fields and fixed a missing cloud runtime dependency.
