# Changelog of SEC EDGAR Full-Text Search Scraper (`logiover/sec-edgar-fulltext-scraper`) Actor

- **URL**: https://apify.com/logiover/sec-edgar-fulltext-scraper/changelog.md
- **Full Actor documentation**: https://apify.com/logiover/sec-edgar-fulltext-scraper.md

## Changelog

### 2026-09-01

- Fleet-wide health check. Verified against this Actor's real run history: 30-day success rate, output row counts, per-field fill rates, and peak memory against the configured memory limit.
- Reviewed for the failure patterns that have cost this fleet runs — unguarded proxy setup, retry loops that can outlast the run's own time budget, and full-page HTML parsing that can exhaust a small container.
- No change to input, output fields or scraping logic.

### 2026-08-11

- Maintenance release: refreshed the build and dependencies.
- Re-verified live execution, non-empty structured output and dataset field/type integrity.
- Reviewed reliability (retries, pagination) and output quality as part of a full-fleet QA pass.

### 2026-08-01

- Added stable run-wide accession-number deduplication. SEC EFTS can return several matching documents from the same filing; the Actor now retains the first hit in upstream order and continues pagination until `maxResults` unique filings are collected or the bounded source window/time budget is exhausted.
- Clarified in the input and dataset contracts that `maxResults` counts unique filings and `fileUrl` identifies the retained matched document. Query, form, date and sort semantics are unchanged.
- Corrected README coverage, direct-proxy default, numeric example and EFTS query-language claims against the current SEC documentation.
- Completed the August 2026 full health check: verified empty/programmatic default, Console UI default, and two source-informed alternative inputs on Apify.
- Confirmed successful live execution, non-empty structured output, dataset-field/type integrity, and logical sample quality within the 5-minute quality window.
- Added stable run-wide accession-number deduplication for SEC EFTS document-level hits, retaining the first upstream-ranked document and applying maxResults to unique filings without changing input query/filter semantics.
- Aligned README, input and dataset contracts with the unique-filing output and retained matched-document URL semantics.
- Corrected README coverage, direct-proxy default, numeric example and query-language claims against current SEC documentation.

### 2026-07-18

- Maintenance & reliability pass — re-verified end-to-end against live data via the Apify API; confirmed the Actor completes successfully and returns non-empty, well-formed results within the 5-minute quality window on its default input.
- Refreshed build so the latest version is current; no change to inputs, output fields or scraping logic.

### 2026-07-12

- **No required inputs** — empty input `{}` now returns filings for a sensible default broad query (recent, relevance-ranked) instead of exiting with nothing. Previously the run required a `query` and produced 0 items when it was missing.
- **Form Types is now a dropdown** — pick one or more SEC form types (10-K, 8-K, S-1, DEF 14A, 13F-HR, Form 4, etc.) from a multi-select list with human-readable labels, instead of typing raw codes. Leave empty to search all forms.
- **New `sort` dropdown** — order results by Relevance (default), Newest first, or Oldest first (by filing date).
- **`relevanceScore` is a real number** — was previously stored as a string.
- **Proxy** — defaults to Apify Proxy (AUTO); a fresh IP is used per page and it automatically falls back to a direct connection on the final retry (EDGAR is an open government server).
- **Graceful runs** — added a ~4 minute time budget; the run pushes what it has and exits successfully.
- **Schema** — added `fields` (with nullable-safe types) to the dataset schema so columns render with correct types and null values are not rejected.
