Skip to content

Validation ceiling: scenario-design gate, computed calibration artifacts, honest multiplier - #5

Merged
vahid-ahmadi merged 1 commit into
mainfrom
validation-calibration-2026-08-05
Aug 5, 2026
Merged

Validation ceiling: scenario-design gate, computed calibration artifacts, honest multiplier#5
vahid-ahmadi merged 1 commit into
mainfrom
validation-calibration-2026-08-05

Conversation

@vahid-ahmadi

Copy link
Copy Markdown
Contributor

Closes the remaining VALIDATION.md targets 2/3 at the achievable ceiling and turns calibration prose into computed, regression-tested artifacts.

Published-figure search (definitive): the v1.1 manual publishes no scenario results (baseline Table 4 + scenario design parameters only); the George/Dafermos SSRN paper (6541398) is paywalled with no open mirror and its scenario set predates v1.1's regulation policies. The open FMM 2023 conference version provides two coarse text anchors, used with wide vintage tolerances.

What is now gated:

  • Scenario design integrity: exact policy-switch sets per scenario vs the same folder's baseline (FF_BAN + ban timings; POWER_SUB alone; GVT_INVEST+GREEN_BONDS+GREEN_POWER for GPI; HOUSING_SUB_RATE=0.4 pinned).
  • FMM 2023 anchors: GPI GDP peak +0.92% (published ≈+1%); baseline 2030 emissions 342.8 (published 'just under 350').
  • Calibration: validation/baseline_vs_external.csv computed by define_uk.validation.baseline_calibration() and drift-tested.
  • Multiplier: our own 1.78 (cum ΔGDP/ΔSPEND_GVT, nominal, full simulation) with the upstream table documented as unusable (M_Impact −32.22 identical across all 8 scenarios; ΔG=104.22 reproducible from no variable — though its ΔGDP=250.76 reconciles with our real-GDP cumulative 250.6).

Suite: 130 passed, 1 skipped.

🤖 Generated with Claude Code

…n, own multiplier

- tests/test_scenario_design.py: each scenario toggles exactly the policy
  switches its published description claims (pinned flag sets incl. the
  40% housing subsidy rate), plus two coarse FMM-2023 anchors (GPI GDP
  peak +0.92% vs ~+1%; baseline 2030 emissions 342.8 vs 'just under 350').
- define_uk/validation.py + validation/baseline_vs_external.csv: the
  external-comparator table is now computed from the cached run and
  regression-tested, not prose.
- Own GPI multiplier 1.78 (cum dGDP / cum dSPEND_GVT, nominal) replaces
  the unusable upstream Multiplier_Summary.csv (scenario-invariant
  columns; TotalDeltaG matches no variable).
- Established and documented: no machine-readable numeric scenario
  results exist for v1.1 (manual stops at baseline; SSRN paper paywalled,
  earlier scenario set), so targets 2/3 are CLOSED at the achievable
  ceiling; sources catalogued in validation/published_targets.json.

Suite: 130 passed, 1 skipped.

Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
@vahid-ahmadi
vahid-ahmadi force-pushed the validation-calibration-2026-08-05 branch from 3cfbe13 to 9e421c1 Compare August 5, 2026 12:02
@vahid-ahmadi
vahid-ahmadi merged commit 27958bd into main Aug 5, 2026
1 check passed
@vahid-ahmadi
vahid-ahmadi deleted the validation-calibration-2026-08-05 branch August 5, 2026 12:03
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant