# Changelog of Website Content Crawler Lite (`fetch_cat/website-content-crawler-lite`) Actor

- **URL**: https://apify.com/fetch\_cat/website-content-crawler-lite/changelog.md
- **Full Actor documentation**: https://apify.com/fetch\_cat/website-content-crawler-lite.md

## Changelog

- 2026-09-25 - Improved the input form by keeping everyday controls visible and grouping specialized settings under Advanced options.

### 0.4

- Fixed Apify proxy execution by replacing invalid base64/hyphenated session IDs with deterministic SHA-256 identifiers using only allowed characters.
- Kept direct access as the default and added proxy setup validation before crawl work and result charging.
- Bounded robots, retry backoff, and page requests by the shared graceful deadline.
- Added injected lifecycle tests for direct and proxy routes, robots.txt, retries, mixed page outcomes, deadlines, and charge-after-save behavior.
- Added fingerprinted pending-work checkpoints and dataset reconciliation so deadline-stopped runs can resume without duplicating stored URLs.
- Enforced a 30-second shutdown reserve before the 300-second platform timeout.

### 0.3

- Added bounded concurrency, transient retry/backoff with `Retry-After`, optional proxy support, and graceful pre-timeout stopping.
- Added response, content, and link limits to prevent oversized pages from exhausting memory or exceeding dataset limits.
- Added redirect/private-network protection, block-page detection, multi-article extraction, and `RUN_SUMMARY`.
- Fixed zero-result and fatal-error runs incorrectly appearing successful.
- Coupled successful page storage to the `page` charge while keeping diagnostic rows free.
- Improved support for documentation sites that safely redirect their public starting URL to a new domain.

### 0.2

- Improved examples, documentation, and reliability.

### 0.1

- Initial version with configurable inputs and structured results.
