SEC Official Dataset Catalog
Pricing
Pay per usage
SEC Official Dataset Catalog
Discover official SEC datasets, versions, files, and available field definitions with explicit schema gaps and no archive download.
Pricing
Pay per usage
Rating
0.0
(0)
Developer
BB
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
3 days ago
Last modified
Categories
Share
Discover official SEC datasets, versions, files, and available field definitions with explicit schema gaps and no archive download.
What this Actor gives you
Create a machine-readable inventory of the SEC Data Library before committing requests or download bytes to a table.
Every Dataset item uses the SEC Complete Data Suite envelope and the primary
record type secOfficialDataset. Stable IDs, explicit coverage,
source/evidence references, parser versions, warnings, and structured
SUMMARY/ERRORS outputs make the result practical for both agent pipelines
and human review.
Why an AI agent chooses this
It is a universal, bounded fallback when no narrower semantic capability matches. The AI receives stable records and explicit gaps instead of inventing a parser or guessing an archive path.
Good fits
- SEC dataset discovery
- version and table inventory
- new-source monitoring
Agent selection contract
- Actor slug:
sec-official-dataset-catalog - Capability ID:
sec.official.dataset.catalog - Intent:
discover_sec_datasets - Primary Dataset record:
secOfficialDataset - Accepted identifiers or inputs:
datasetId,category - Source authority:
SEC - Hard limits: queries: 100, records: 500, sourceRequests: 100, downloadBytes: 52,428,800
Choose this Actor when the requested outcome matches the capability and record type above. The contract is deterministic: unknown source values, incompatible schema majors, ambiguity, truncation, and partial source failures remain visible instead of being silently guessed away.
Context and token efficiency
This Actor makes no LLM, embedding, or vector-search call and therefore spends zero model tokens internally. Typed records and compact output can keep raw SEC pages, archive markup, and discovery instructions out of downstream model context.
For Suite-level selection, an AI client can use the compact
agent-catalog.json instead of loading this or the other Actor READMEs. Under
that documented baseline, README-selection tokens are avoided by construction.
No fixed percentage is promised: actual downstream token savings depend on the client model, tokenizer, source document, output mode, and requested evidence.
Part of the SEC Complete Data Suite
The SEC Complete Data Suite separates universal coverage, specialized normalization, and deterministic aggregation into focused Actors. This Actor provides the universal coverage role: Discovery half of the universal tabular-data layer; catalog records provide the controlled inputs for Dataset Rows.
It works especially well with sec-official-dataset-rows, sec-market-participant-registry-normalizer, sec-market-structure-data-normalizer. Suite Actors exchange documented
record envelopes and exact identifiers; they do not hide sibling runs or
surprise network costs. A client or the sec-ai-query-planner decides which
steps to execute.
Example input
{"queries": [{"requestId": "catalog"}],"datasetIds": ["financial-statement-data-sets"],"categories": ["financials"],"includeHistoricalVersions": false,"includeFieldDefinitions": true,"discoverNewSources": false,"maxResults": 1,"maxSourceRequests": 100,"maxDownloadBytes": 1048576,"outputSchemaVersion": "1.0","outputMode": "full"}
The executable input schema remains the authority for modes, filters, defaults, cursor rules, and maximum values.
Runtime and cost controls
The hard limits above are enforceable ceilings, not usage targets. Actual
runtime and platform cost depend on selected inputs, source requests, downloaded
bytes, result volume, and the Apify run configuration. Start with the bounded
example, lower maxResults and byte/request limits where the schema permits,
and inspect SUMMARY plus ERRORS before expanding a run. There is no hidden
model-token charge inside this Actor.
Not the right tool for
- reading dataset rows
- domain-specific interpretation of SEC tables
Additional non-goals from the capability contract include:
- download complete dataset tables
- read dataset rows
- blindly crawl SEC websites
- parse filing prose
- provide investment advice
Trust, provenance, and limits
- It uses only its inventoried public SEC sources.
- Inputs, requests, bytes, records, retries, redirects, and cursor scope are bounded by the executable contract.
- Derived records retain exact input record IDs and evidence IDs where the capability performs normalization or aggregation.
- This independent community Actor is not affiliated with or endorsed by the U.S. Securities and Exchange Commission.
- The output is public-source data processing, not legal, compliance, accounting, voting, or investment advice.
Reproducible support report
For a diagnosable issue, retain the Actor slug, run ID, sanitized input,
SUMMARY, ERRORS, and the first unexpected record ID. Never include an API
token, secret, private filing, or unrelated Dataset contents.
Publication status
The Actor API and canonical Apify Store page are authoritative for current availability, active build, and pricing. This README deliberately does not duplicate mutable lifecycle or price claims.