# Changelog of Zenodo Scraper — Research Datasets, Papers & Software (`logiover/zenodo-scraper`) Actor

- **URL**: https://apify.com/logiover/zenodo-scraper/changelog.md
- **Full Actor documentation**: https://apify.com/logiover/zenodo-scraper.md

## Changelog

### 2026-09-01

- Fleet-wide health check. Verified against this Actor's real run history: 30-day success rate, output row counts, per-field fill rates, and peak memory against the configured memory limit.
- Reviewed for the failure patterns that have cost this fleet runs — unguarded proxy setup, retry loops that can outlast the run's own time budget, and full-page HTML parsing that can exhaust a small container.
- No change to input, output fields or scraping logic.

### 2026-08-11

- Maintenance release: refreshed the build and dependencies.
- Re-verified live execution, non-empty structured output and dataset field/type integrity.
- Reviewed reliability (retries, pagination) and output quality as part of a full-fleet QA pass.

### 2026-08-01

- Completed the August 2026 full health check: verified empty/programmatic default, Console UI default, and two source-informed alternative inputs on Apify.
- Confirmed successful live execution, non-empty structured output, dataset-field/type integrity, and logical sample quality within the 5-minute quality window.
- Added stable run-wide ID/DOI/canonical-URL deduplication across pagination and record-detail inputs, plus repeated next-page URL protection, so overlapping Zenodo pages cannot emit duplicate records.

### 2026-07-18

- Maintenance & reliability pass — re-verified end-to-end against live data via the Apify API; confirmed the Actor completes successfully and returns non-empty, well-formed results within the 5-minute quality window on its default input.
- Refreshed build so the latest version is current; no change to inputs, output fields or scraping logic.

### 2026-07-12

- Empty input `{}` now returns data: with no query the actor browses the most recent Zenodo records (bounded by `maxResults`, default 200) instead of exiting.
- Removed the mandatory `mode` requirement — every field is optional.
- Added a **Sort Order** dropdown (Auto / Most recent / Best match / Oldest / Most viewed / Least viewed). "Auto" picks most-recent when browsing and best-match when a query is supplied.
- Expanded the **Resource Type** dropdown with more Zenodo types (poster, presentation, physical object, workflow, …).
- `recordDetail`/`communityRecords` modes now fall back to a recent-records browse when their required field is missing, instead of erroring.
- Proxy is now actually wired in: uses Apify Proxy (AUTO) with a **direct-connection fallback** on the final retry (Zenodo is a clean public API). No forced DATACENTER group.
- `views` and `downloads` are now guaranteed to be stored as real numbers.
- Added a pagination **time budget** (~4 min): the run pushes what it has collected and exits successfully rather than running long.
- Top-level errors are logged (not hard-thrown) so a partial run still succeeds.
- Added `fields` metadata to the dataset schema (titles + descriptions) for a nicer Dataset view.
