Forecast evaluation
Four selection folds, frozen baselines, 12-level WRMSSE, and one aggregate-only official evaluation.
Evidence map
Every capability below points to a frozen artifact, a receipt, or a testable control. Status is explicit where publication still depends on later release points.
01 · Capability evidence
Four selection folds, frozen baselines, 12-level WRMSSE, and one aggregate-only official evaluation.
Six predeclared asymmetric-cost ratios and leakage-safe quantile recalibration at bottom level before aggregation.
Twenty-seven simulated policies with lost sales, empirical buffers, denominator rules, and aggregate-only exposure outputs.
Forecast-to-actual PVM, linked three-statement Excel model, revolver logic, scenarios, sensitivity, and visible checks.
Native four-page semantic model with governed measures, hierarchy drill, refresh evidence, PDF export, and contract tests.
Canonical-period selection, concept mapping, filing lineage, derived free cash flow, ratios, and reconciliation checks.
SCD Type 2 price history, checksum gates, hierarchy controls, immutable receipts, and fail-closed lifecycles.
Four frozen FRED observations used only as separate descriptive context, with no regime or causal claim.
Pinned Python graph, portable web build, content-addressed evidence, focused tests, and explicit rerun boundaries.
02 · Release boundary