# Changelog of Openverse Scraper — CC Images, Audio & Media API (`logiover/openverse-scraper`) Actor

- **URL**: https://apify.com/logiover/openverse-scraper/changelog.md
- **Full Actor documentation**: https://apify.com/logiover/openverse-scraper.md

## Changelog

### 2026-09-23

- Fleet-wide quality audit. Verified end to end against live data and re-checked the input schema, the output columns and the run configuration.
- Run reliability reviewed: 100.0% of public runs succeeded in the last 30 days.
- Input schema, output schema and pricing configuration reviewed.

### 2026-09-01

- Fleet-wide health check. Verified against this Actor's real run history: 30-day success rate, output row counts, per-field fill rates, and peak memory against the configured memory limit.
- Reviewed for the failure patterns that have cost this fleet runs — unguarded proxy setup, retry loops that can outlast the run's own time budget, and full-page HTML parsing that can exhaust a small container.
- No change to input, output fields or scraping logic.

### 2026-08-11

- Maintenance release: refreshed the build and dependencies.
- Re-verified live execution, non-empty structured output and dataset field/type integrity.
- Reviewed reliability (retries, pagination) and output quality as part of a full-fleet QA pass.

### 2026-08-01

- Completed the August 2026 full health check: verified empty/programmatic default, Console UI default, and two source-informed alternative inputs on Apify.
- Confirmed successful live execution, non-empty structured output, dataset-field/type integrity, and logical sample quality within the 5-minute quality window.

### 2026-07-18

- Maintenance & reliability pass — re-verified end-to-end against live data via the Apify API; confirmed the Actor completes successfully and returns non-empty, well-formed results within the 5-minute quality window on its default input.
- Refreshed build so the latest version is current; no change to inputs, output fields or scraping logic.

### 2026-07-12

- **No query required.** Empty input `{}` now returns data: an empty search browses a broad default listing of popular/recent Creative Commons media (images by default) instead of exiting. `query` is fully optional.
- **Dropdowns for finite filters.** `license`, `licenseType` (usage), `source`/provider, and `extension`/format are now select menus with human-readable labels; `query` and `imageId` stay free-text.
- **Numbers as real numbers.** `width`, `height`, and the new `fileSize`, `duration`, `bitRate`, `sampleRate` are stored as numeric values (previously width/height were strings). Audio-only numeric fields are added and null for images.
- **New optional filters.** Added `licenseType` (commercial / modification / all-cc) and `extension` (file format) filters.
- **Proxy.** Defaults to Apify automatic proxy with a direct-connection fallback on request failure (public API); no forced datacenter group.
- **Graceful runs.** Added a ~4-minute pagination time budget; the top level logs errors instead of hard-failing the run.
- **Schemas.** Added a nullable-safe dataset `fields` schema (title + description per field); dataset view renders numeric columns with `format: number`. Nothing is required.
