SEC Official Dataset Rows avatar

SEC Official Dataset Rows

Pricing

from $0.05 / 1,000 dataset rows

Go to Apify Store
SEC Official Dataset Rows

SEC Official Dataset Rows

Read bounded, typed rows from catalogued public SEC dataset tables while retaining original columns and source provenance.

Pricing

from $0.05 / 1,000 dataset rows

Rating

0.0

(0)

Developer

BB

BB

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

15 days ago

Last modified

Share

Read bounded, typed rows from catalogued public SEC dataset tables while retaining original columns and source provenance.

What this Actor gives you

Stream selected rows from a known SEC dataset version without silently converting the source into an unrelated domain model.

Every Dataset item uses the SEC Complete Data Suite envelope and the primary record type secOfficialDatasetRow. Stable IDs, explicit coverage, source/evidence references, parser versions, warnings, and structured SUMMARY/ERRORS outputs make the result practical for both agent pipelines and human review.

Why an AI agent chooses this

It is a universal, bounded fallback when no narrower semantic capability matches. The AI receives stable records and explicit gaps instead of inventing a parser or guessing an archive path.

Good fits

  • official SEC bulk-table access
  • bounded dataset extraction
  • generic fallback for newly published datasets

Agent selection contract

  • Actor slug: sec-official-dataset-rows
  • Capability ID: sec.official.dataset.rows
  • Intent: read_sec_dataset_rows
  • Primary Dataset record: secOfficialDatasetRow
  • Accepted identifiers or inputs: datasetVersionId+table
  • Source authority: SEC
  • Hard limits: queries: 20, records: 50,000, sourceRequests: 1,000, downloadBytes: 262,144,000

Choose this Actor when the requested outcome matches the capability and record type above. The contract is deterministic: unknown source values, incompatible schema majors, ambiguity, truncation, and partial source failures remain visible instead of being silently guessed away.

Context and token efficiency

This Actor makes no LLM, embedding, or vector-search call and therefore spends zero model tokens internally. Typed records and compact output can keep raw SEC pages, archive markup, and discovery instructions out of downstream model context.

For Suite-level selection, an AI client can use the compact agent-catalog.json instead of loading this or the other Actor READMEs. Under that documented baseline, README-selection tokens are avoided by construction.

No fixed percentage is promised: actual downstream token savings depend on the client model, tokenizer, source document, output mode, and requested evidence.

Part of the SEC Complete Data Suite

The SEC Complete Data Suite separates universal coverage, specialized normalization, and deterministic aggregation into focused Actors. This Actor provides the universal coverage role: Universal row-access layer and fallback source for dataset-backed specialty Actors.

It works especially well with sec-official-dataset-catalog, sec-market-structure-data-normalizer, sec-registered-product-normalizer. Suite Actors exchange documented record envelopes and exact identifiers; they do not hide sibling runs or surprise network costs. A client or the sec-ai-query-planner decides which steps to execute.

Example input

{
"queries": [
{
"requestId": "financials-sub",
"datasetVersionId": "financial-statement-data-sets:2026-03-31",
"table": "sub",
"columns": [
"adsh",
"cik",
"form",
"filed"
],
"includeAllColumns": false,
"filters": [
{
"field": "cik",
"operator": "eq",
"value": "320193"
}
]
}
],
"maxResults": 3,
"maxSourceRequests": 1,
"maxDownloadBytes": 104857600,
"maxCompressedBytes": 104857600,
"maxUncompressedBytes": 536870912,
"outputSchemaVersion": "1.0",
"outputMode": "full"
}

The executable input schema remains the authority for modes, filters, defaults, cursor rules, and maximum values.

Runtime and cost controls

The hard limits above are enforceable ceilings, not usage targets. Actual runtime and platform cost depend on selected inputs, source requests, downloaded bytes, result volume, and the Apify run configuration. Start with the bounded example, lower maxResults and byte/request limits where the schema permits, and inspect SUMMARY plus ERRORS before expanding a run. There is no hidden model-token charge inside this Actor.

Not the right tool for

  • uncatalogued downloads or arbitrary URLs
  • semantic claims beyond the source columns

Additional non-goals from the capability contract include:

  • interpret dataset rows as domain-specific facts
  • accept arbitrary user URLs
  • download uncatalogued artifacts
  • run sibling actors
  • parse filing prose

Trust, provenance, and limits

  • It uses only its inventoried public SEC sources.
  • Inputs, requests, bytes, records, retries, redirects, and cursor scope are bounded by the executable contract.
  • Derived records retain exact input record IDs and evidence IDs where the capability performs normalization or aggregation.
  • This independent community Actor is not affiliated with or endorsed by the U.S. Securities and Exchange Commission.
  • The output is public-source data processing, not legal, compliance, accounting, voting, or investment advice.

Reproducible support report

For a diagnosable issue, retain the Actor slug, run ID, sanitized input, SUMMARY, ERRORS, and the first unexpected record ID. Never include an API token, secret, private filing, or unrelated Dataset contents.

Publication status

The Actor API and canonical Apify Store page are authoritative for current availability, active build, and pricing. This README deliberately does not duplicate mutable lifecycle or price claims.