# Changelog of Wikipedia Category Scraper — Article Lists & Data (`logiover/wikipedia-category-scraper`) Actor

- **URL**: https://apify.com/logiover/wikipedia-category-scraper/changelog.md
- **Full Actor documentation**: https://apify.com/logiover/wikipedia-category-scraper.md

## Changelog

### 2026-09-01

- Fleet-wide health check. Verified against this Actor's real run history: 30-day success rate, output row counts, per-field fill rates, and peak memory against the configured memory limit.
- Reviewed for the failure patterns that have cost this fleet runs — unguarded proxy setup, retry loops that can outlast the run's own time budget, and full-page HTML parsing that can exhaust a small container.
- No change to input, output fields or scraping logic.

### 2026-08-11

- Maintenance release: refreshed the build and dependencies.
- Re-verified live execution, non-empty structured output and dataset field/type integrity.
- Reviewed reliability (retries, pagination) and output quality as part of a full-fleet QA pass.

### 2026-08-01

- Corrected the German health-check category to its localized `Programmiersprache` name and verified 30 valid German Wikipedia article rows.
- Completed the August 2026 full health check: verified empty/programmatic default, Console UI default, and two source-informed alternative inputs on Apify.
- Confirmed successful live execution, non-empty structured output, dataset-field/type integrity, and logical sample quality within the 5-minute quality window.
- Corrected the German variation from the English Programming\_languages slug to the localized Programmiersprache category; the exact input produced 30 German article rows with valid localized URLs.
- Removed the README's pageId claim and sample field because neither category-list nor summary mode emits a pageId.

### 2026-07-18

- Maintenance & reliability pass — re-verified end-to-end against live data via the Apify API; confirmed the Actor completes successfully and returns non-empty, well-formed results within the 5-minute quality window on its default input.
- Refreshed build so the latest version is current; no change to inputs, output fields or scraping logic.

### 2026-07-12

- **Runs with empty input** — all fields are optional (`required: []`). With no input, the actor now scrapes a popular default category ("Large\_language\_models" on English Wikipedia) instead of exiting empty. Bounded by `maxArticles`.
- **More languages** — the language dropdown now covers 20 Wikipedia editions (added Dutch, Polish, Swedish, Ukrainian, Turkish, Persian, Korean, Indonesian, Vietnamese, Hebrew).
- **Cleaner input** — Category Names is now the primary field; full Category URLs are kept as an advanced option. Category names accept an optional `Category:` prefix.
- **Robustness** — added a ~4-minute time budget (stops queueing new pages and exits gracefully with collected data), a failed-request handler, and top-level error handling so the run does not hard-fail.
- **Schemas** — added a documented dataset `fields` schema (title + description per field) and added the language column to the dataset view. Output field keys are unchanged.
