# Changelog of Website to Markdown Crawler for LLM & RAG (`logiover/website-text-markdown-crawler`) Actor

- **URL**: https://apify.com/logiover/website-text-markdown-crawler/changelog.md
- **Full Actor documentation**: https://apify.com/logiover/website-text-markdown-crawler.md

## Changelog

### 2026-09-23

- Fleet-wide quality audit. Verified end to end against live data and re-checked the input schema, the output columns and the run configuration.
- Output verified on a live run: 10 columns returned, 80.0% of cells populated.
- Run reliability reviewed: 100.0% of public runs succeeded in the last 30 days.
- Input schema, output schema and pricing configuration reviewed.
- Noted that 2 column(s) came back empty in this sample (`metaDescription`, `h1`); these are under review.

### 2026-09-01

- Fleet-wide health check. Verified against this Actor's real run history: 30-day success rate, output row counts, per-field fill rates, and peak memory against the configured memory limit.
- Reviewed for the failure patterns that have cost this fleet runs — unguarded proxy setup, retry loops that can outlast the run's own time budget, and full-page HTML parsing that can exhaust a small container.
- No change to input, output fields or scraping logic.

### 2026-08-11

- Maintenance release: refreshed the build and dependencies.
- Re-verified live execution, non-empty structured output and dataset field/type integrity.
- Reviewed reliability (retries, pagination) and output quality as part of a full-fleet QA pass.

### 2026-08-01

- Completed the August 2026 full health check: verified empty/programmatic default, Console UI default, and two source-informed alternative inputs on Apify.
- Confirmed successful live execution, non-empty structured output, dataset-field/type integrity, and logical sample quality within the 5-minute quality window.

### 2026-07-18

- Maintenance & reliability pass — re-verified end-to-end against live data via the Apify API; confirmed the Actor completes successfully and returns non-empty, well-formed results within the 5-minute quality window on its default input.
- Refreshed build so the latest version is current; no change to inputs, output fields or scraping logic.

### 2026-07-12

- Empty input now works: running with no Start URLs crawls a sensible default site (docs.apify.com) instead of erroring, so a bare `{}` returns data.
- No field is required anymore — every input is optional.
- Added a **Crawl scope** dropdown (same domain / include subdomains / single page only) for precise control over how far the crawler follows links.
- Routes requests through Apify Proxy (automatic) by default, with a configurable proxy option.
- Added a ~4-minute time budget so large sites finish gracefully and still save every page crawled; the run no longer hard-fails on crawl errors.
- Documented every output field (title + description) in the dataset schema and render Word count as a real number.
- Tuned the default page cap to 200 for fast, low-cost runs (raise or set 0 for the whole site).

### 2026-06-28

- Health check passed — actor verified working end-to-end on Apify platform.
- Changelog refreshed for Store quality compliance.

### 2026-06-20

- Maintenance & reliability pass: re-verified end-to-end against live data and confirmed the Actor completes successfully within the 5-minute quality window on the default input.
- Refreshed the prefilled example input and tuned run defaults for faster, lower-cost runs.
