Skip to content
Fiscal Receipts

Research tools / Analyze

Ask the data

Inspect a dataset, run a prepared query, or write your own SQL. Work from the same snapshot that powers the program pages.

Browse and query all 16 Fiscal Receipts datasets using SQL. Queries run entirely in your browser — no server receives your SQL or sees any intermediate results. Powered by DuckDB-WASM; parquet files are streamed on demand (HTTP range requests).

Which programs are included in this snapshot?

Corpus: 2,562 browsable program pages; 1,936 of them carry detail-grade R-2/P-40 J-book data — excludes personnel, O&M, and R-1/P-1 lines that lack R-2/P-40 project detail. what separates the two tiers? →

Queries run entirely in your browser — no data leaves your machine. The DuckDB-WASM engine loads on demand (~4 MB).

One row per (program element × amount type) figure in the FY2026 President's Budget R-1/P-1 workbooks, carrying the source sheet and cell address it was read from. Fenced to the PB2026 edition. Add rows only: the P-1 flags each row Add or Non-Add, and Non-Add rows are memo lines the exhibit does not add into its own totals, so summing the workbook without that filter double-counts. P-1R rows are the reserve-component subset of P-1 lines and are never summed with them.

Canned queries

Dataset inventory

Row counts and sizes are read from the parquet files this build shipped — never authored by hand. Each scope line says what one row of that dataset is, so a row count can be compared against the right denominator.

budget_lines8,549 rows181 KBcitedOne row per (program element × amount type) figure in the FY2026 President's Budget R-1/P-1 workbooks, carrying the source sheet and cell address it was read from. Fenced to the PB2026 edition. Add rows only: the P-1 flags each row Add or Non-Add, and Non-Add rows are memo lines the exhibit does not add into its own totals, so summing the workbook without that filter double-counts. P-1R rows are the reserve-component subset of P-1 lines and are never summed with them.
budget_lines_decade38,855 rows903 KBcitedThe decade sibling of budget_lines: one row per (President's Budget edition × program element × amount type) figure across all ten editions PB2017–PB2026, each cited to its own edition's workbook cell. Editions are parallel publications, never reconciled.
jbook_details21,993 rows379 KBcitedOne row per (program element × project × budget scenario) cost figure extracted from J-book R-2/P-40 XML, with its XML element path and source-PDF SHA-256. `account` is the appropriation the figure was filed under (NULL for R-2/RDT&E rows, which carry none): ten PB2026 budget-line codes are shared by two different Navy programs in two different appropriations, so summing this table by pe_bli alone adds those pairs together.
jbook_narratives11,717 rows3.8 MBcitedOne row per J-book narrative text block (mission, description, justification or accomplishment/planned-program) with its XML element path and source-PDF SHA-256. Fenced to the PB2026 edition, plus the PB2017–PB2025 narratives a program-lineage edge cites: citation targets only, never a program page's own prose.
fct_budget_trajectory1,988 rows32 KBcitedOne row per (program element × organization) with FY2024 actuals, FY2025 total, FY2026 total and the FY25→FY26 delta pivoted side by side.
dim_programs1,936 rows45 KBcitedOne row per program element that has full R-2/P-40 J-book detail (the detail-grade tier, NOT the full page universe; see the corpus statement on /data/), with org, exhibit family, project count and reconciliation status.
dim_entities116,442 rows3.6 MBcitedOne row per contractor entity family, name-normalized across USAspending/SAM.gov UEI registrations, with its UEI count, total DoD obligation and worst match-confidence tier. The sam_* columns carry the SAM.gov Entity Management record of the family's dominant registration (status, CAGE, legal business name, business types, primary NAICS, expiry) where the bounded extract has reached it — enrichment, never an input to the confidence tier, and NULL wherever it has not.
fct_influence210 rows6 KBcitedOne row per (contractor family × filing year) of Senate LDA lobbying totals, beside the family's DoD obligations (repeated on each year row). Income and expense are non-additive: a self-filer's expense can include its outside firms' income, so lobbying_total_usd, a plain sum, can double-count.
fct_program_lobbying12,571 rows159 KBcitedOne row per (LDA filing × matched program element) mention — a filing appears once for every program its issue text has keyword co-occurrence with (an exact PE/BLI code, a curated alias, or >=2 distinct title words; see evidence_kind), never a claim that the filing names the program. Rows count evidence-tiered mentions, not filings.
fct_budget_to_awards3,685 rows63 KBcitedOne row per (program element × USAspending award) crosswalk link, each carrying its match method and confidence tier. A link is an inference, not a reported fact.
dim_geography1,048 rows14 KBcitedOne row per (pop_state × pop_district) place of performance, with transaction count and total obligation over every DoD award transaction the site loads that records a pop_state, not only crosswalk-linked awards. pop_state is not normalized (contracts: two-letter code; assistance: state name).
fct_district_totals191 rows3 KBcitedOne row per (state × congressional district), with obligation dollars counted once per DISTINCT high-confidence-crosswalked award — the district headline. Its per-(district, program element, appropriation account) sibling, fct_district_programs, is not summable: an award matched to N program elements appears N times with the same dollars there.
fct_program_concentration536 rows18 KBcitedOne row per program element carrying at least one published crosswalk link, on TWO bases: *_all over every published link (high and medium confidence), *_high over high-confidence links alone. hhi_high and top_family_high are NULL below the floor: 3 linked awards, 2 families holding positive dollars (positive_family_count_high), positive program_dollars_high (ROADMAP #80).
fct_state_per_capita6 rows4 KBcitedOne row per (jurisdiction × comparable spending category × fiscal year) in the CA/CT state-checkbook pilot, with population and per-capita amount.
fct_improper_exposure15 rows1 KBcitedOne row per federal agency reporting improper payments, with the derived dollar exposure and weighted error rate from paymentaccuracy.gov.
dim_lobbyists1,612 rows100 KBcitedOne row per individual lobbyist named in Senate LDA filings, with covered-position and revolving-door flags and the UUID/URL of the filing that disclosed them.

Built: September 26, 2026. Citation methodology: see methodology.

California data covers FY2025 only — CA Open Fi$Cal updates on a lag; prior years not yet ingested. why FY2025 only? →