SEC Official Dataset Rows
Pricing
from $0.05 / 1,000 dataset rows
SEC Official Dataset Rows
Read bounded, typed rows from catalogued public SEC dataset tables while retaining original columns and source provenance.
Pricing
from $0.05 / 1,000 dataset rows
Rating
0.0
(0)
Developer
BB
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
15 days ago
Last modified
Categories
Share
Read bounded, typed rows from catalogued public SEC dataset tables while retaining original columns and source provenance.
What this Actor gives you
Stream selected rows from a known SEC dataset version without silently converting the source into an unrelated domain model.
Every Dataset item uses the SEC Complete Data Suite envelope and the primary
record type secOfficialDatasetRow. Stable IDs, explicit coverage,
source/evidence references, parser versions, warnings, and structured
SUMMARY/ERRORS outputs make the result practical for both agent pipelines
and human review.
Why an AI agent chooses this
It is a universal, bounded fallback when no narrower semantic capability matches. The AI receives stable records and explicit gaps instead of inventing a parser or guessing an archive path.
Good fits
- official SEC bulk-table access
- bounded dataset extraction
- generic fallback for newly published datasets
Agent selection contract
- Actor slug:
sec-official-dataset-rows - Capability ID:
sec.official.dataset.rows - Intent:
read_sec_dataset_rows - Primary Dataset record:
secOfficialDatasetRow - Accepted identifiers or inputs:
datasetVersionId+table - Source authority:
SEC - Hard limits: queries: 20, records: 50,000, sourceRequests: 1,000, downloadBytes: 262,144,000
Choose this Actor when the requested outcome matches the capability and record type above. The contract is deterministic: unknown source values, incompatible schema majors, ambiguity, truncation, and partial source failures remain visible instead of being silently guessed away.
Context and token efficiency
This Actor makes no LLM, embedding, or vector-search call and therefore spends zero model tokens internally. Typed records and compact output can keep raw SEC pages, archive markup, and discovery instructions out of downstream model context.
For Suite-level selection, an AI client can use the compact
agent-catalog.json instead of loading this or the other Actor READMEs. Under
that documented baseline, README-selection tokens are avoided by construction.
No fixed percentage is promised: actual downstream token savings depend on the client model, tokenizer, source document, output mode, and requested evidence.
Part of the SEC Complete Data Suite
The SEC Complete Data Suite separates universal coverage, specialized normalization, and deterministic aggregation into focused Actors. This Actor provides the universal coverage role: Universal row-access layer and fallback source for dataset-backed specialty Actors.
It works especially well with sec-official-dataset-catalog, sec-market-structure-data-normalizer, sec-registered-product-normalizer. Suite Actors exchange documented
record envelopes and exact identifiers; they do not hide sibling runs or
surprise network costs. A client or the sec-ai-query-planner decides which
steps to execute.
Example input
{"queries": [{"requestId": "financials-sub","datasetVersionId": "financial-statement-data-sets:2026-03-31","table": "sub","columns": ["adsh","cik","form","filed"],"includeAllColumns": false,"filters": [{"field": "cik","operator": "eq","value": "320193"}]}],"maxResults": 3,"maxSourceRequests": 1,"maxDownloadBytes": 104857600,"maxCompressedBytes": 104857600,"maxUncompressedBytes": 536870912,"outputSchemaVersion": "1.0","outputMode": "full"}
The executable input schema remains the authority for modes, filters, defaults, cursor rules, and maximum values.
Runtime and cost controls
The hard limits above are enforceable ceilings, not usage targets. Actual
runtime and platform cost depend on selected inputs, source requests, downloaded
bytes, result volume, and the Apify run configuration. Start with the bounded
example, lower maxResults and byte/request limits where the schema permits,
and inspect SUMMARY plus ERRORS before expanding a run. There is no hidden
model-token charge inside this Actor.
Not the right tool for
- uncatalogued downloads or arbitrary URLs
- semantic claims beyond the source columns
Additional non-goals from the capability contract include:
- interpret dataset rows as domain-specific facts
- accept arbitrary user URLs
- download uncatalogued artifacts
- run sibling actors
- parse filing prose
Trust, provenance, and limits
- It uses only its inventoried public SEC sources.
- Inputs, requests, bytes, records, retries, redirects, and cursor scope are bounded by the executable contract.
- Derived records retain exact input record IDs and evidence IDs where the capability performs normalization or aggregation.
- This independent community Actor is not affiliated with or endorsed by the U.S. Securities and Exchange Commission.
- The output is public-source data processing, not legal, compliance, accounting, voting, or investment advice.
Reproducible support report
For a diagnosable issue, retain the Actor slug, run ID, sanitized input,
SUMMARY, ERRORS, and the first unexpected record ID. Never include an API
token, secret, private filing, or unrelated Dataset contents.
Publication status
The Actor API and canonical Apify Store page are authoritative for current availability, active build, and pricing. This README deliberately does not duplicate mutable lifecycle or price claims.