Tracks the March 2026 EFO to 0.15% GDP MAPE with published anchors; free-running error is 5.75%.
Open economic models · public evidence
Open UK and US economic models, scored in public.
Run the models yourself. Every result carries its assumptions, official data vintage, and validation — and the UK forecasts are scored in public.
One path from economic question to checkable result.
Score a tax or benefit change
The microsimulation computes the direct cost and who pays; the OBR emulator returns the economy-wide feedback — one reform, scored end to end.
score a reform →Test a policy or shock
Choose among seven open household and macro models, then inspect each model’s assumptions, evidence, and limits.
choose a macro model →See what changed
Follow activity, prices, labour, rates, upcoming releases, and the official-data vintages behind every figure.
see the economy →Check forecasts against what happened
Archive forecast rounds before results arrive, preserve official-data vintages, and score every realised period. A US record follows once a US forecasting model is hosted.
inspect the record →Explore all seven models or compare their assumptions side by side.
Latest outturns beside the model's near-term view.
ONS outturns (as of 2026-07-26) beside archived forecast rounds. Full horizon →
FRED outturns (as of 2026-08-04) beside the FRB/US LONGBASE conditioning baseline — not a forecast. Full sources →
What the evidence shows—and what it doesn't.
Each model is checked against the strongest available benchmark. Replication, forecast performance, statutory correctness, and calibration are different kinds of evidence—and where no independent check exists, we say so. Compare the evidence →
Beats a random-walk baseline on CPI across 49 rolling origins; measured band coverage runs below nominal.
Matches the Fed’s pyfrbus to within the Fed’s own two releases’ disagreement (~1×10⁻⁸); no predictive claim.
Hits every Auclert et al. (2021) calibration target; responses are first-order and stylized.
Rules are tested against statute; population totals inherit survey uncertainty.
Targets are met by construction; no independent outcome benchmark exists.
Baseline macro block replicates the manual; scenario deltas are design-gated and paper-anchored — no published scenario numbers exist. Deltas only, never levels.
Run hosted models—or use the code directly.
Connect the public MCP server with no PolicyEngine account or API key, use the shared CLI, or call each Python package directly. Connect and try a model →