# Changelog of Sitemap to URL Crawler — Extract Sitemap.xml URLs (`logiover/sitemap-to-url-crawler`) Actor

- **URL**: https://apify.com/logiover/sitemap-to-url-crawler/changelog.md
- **Full Actor documentation**: https://apify.com/logiover/sitemap-to-url-crawler.md

## Changelog

### 2026-09-01

- Fleet-wide health check. Verified against this Actor's real run history: 30-day success rate, output row counts, per-field fill rates, and peak memory against the configured memory limit.
- Reviewed for the failure patterns that have cost this fleet runs — unguarded proxy setup, retry loops that can outlast the run's own time budget, and full-page HTML parsing that can exhaust a small container.
- No change to input, output fields or scraping logic.

### 2026-08-11

- Maintenance release: refreshed the build and dependencies.
- Re-verified live execution, non-empty structured output and dataset field/type integrity.
- Reviewed reliability (retries, pagination) and output quality as part of a full-fleet QA pass.

### 2026-08-01

- Completed the August 2026 full health check: verified empty/programmatic default, Console UI default, and two source-informed alternative inputs on Apify.
- Confirmed successful live execution, non-empty structured output, dataset-field/type integrity, and logical sample quality within the 5-minute quality window.
- Fixed the run Output link from `{{links.apiDefaultDatasetUrl}}` to `{{links.apiDefaultDatasetUrl}}/items` so the results table opens the dataset items endpoint.
- Made all-root invalid, blocked, or empty sitemap runs fail explicitly and replaced the retired New York Times sitemap variation with a verified MDN sitemap index.

### 2026-07-18

- Maintenance & reliability pass — re-verified end-to-end against live data via the Apify API; confirmed the Actor completes successfully and returns non-empty, well-formed results within the 5-minute quality window on its default input.
- Refreshed build so the latest version is current; no change to inputs, output fields or scraping logic.

### 2026-07-12

- **Empty input now works.** Running with `{}` (no start URL) now falls back to a large public demo site (apify.com) and returns URLs, instead of finishing with zero results. Add your own `startUrls` to crawl any site.
- **No required fields.** Removed the `startUrls` requirement — every input is now optional.
- **Bounded default.** `maxUrls` keeps its sensible 10,000 default so the demo run is fast and cheap; raise it to pull an entire large site.
- **Proxy:** defaults to Apify Proxy (AUTO) when none is supplied (no forced DATACENTER group).
- **Graceful long runs:** added a ~4.5-minute time budget on huge sitemap indexes — the run stops traversing, keeps everything collected, and exits successfully.
- Added a dataset field schema (titles + descriptions) and surfaced `sourceSitemap` in the default table view. `priority` remains a real number. No existing field keys were renamed.

### 2026-06-28

- Health check passed — actor verified working end-to-end on Apify platform.
- Changelog refreshed for Store quality compliance.

### 2026-06-20

- Maintenance & reliability pass: re-verified end-to-end against live data and confirmed the Actor completes successfully within the 5-minute quality window on the default input.
- Refreshed the prefilled example input and tuned run defaults for faster, lower-cost runs.
