Finance.Foundation
Datasets · seed-0.1 · 2026-08-28

Data

Every dataset on Finance Foundation: what it covers, where it comes from, how it is processed, and how to fetch it as JSON.

The data pipeline

Finance Foundation separates data by processing stage. Raw inputs are never mixed into the presentation layer — every published number has passed through the full pipeline.

raw (public filings, official statistics, primary sources) → normalized (one schema, one unit system, one id per entity) → enriched (relationships: lists_on, domiciled_in, managed_by…) → scored (coverage & confidence flags) → audited (human review against primary sources) → final (published — what you see on this site and in the API)

Current status — honest version

The current dataset is seed-0.1 · 2026-08-28: a hand-audited seed covering the most systemically important entities in each class. Figures (market cap, AUM, GDP, bank assets) are approximate reference values from public sources, rounded, dated Q2 2026, and not real-time. The architecture has a provider-abstraction layer, so live sources (filings, market data, macro series) can be attached per dataset without changing the API contract.

Sources & licensing

Sources

Company figures: public filings and listed-market data. Macro: IMF, national statistics offices, central banks. Market/venue data: exchange operators and ISO 10383 MIC registry. Crypto: public on-chain data. Each entity page carries its provenance line.

Licensing

The Finance Foundation seed dataset and schema are published under CC BY 4.0 — free to use with attribution. We do not redistribute licensed vendor datasets, and identifiers that require licenses (CUSIP, SEDOL) are referenced, not republished.