Skill

Verification

Per-Run Model Skill · The NOAA Headline, Scored

How each model run actually performed. The headline is the one NOAA/EMC live by: Northern-Hemisphere 500 hPa anomaly correlation by lead time — area-weighted over 20–80°N, each forecast scored against its own verifying analysis, anomalies taken vs the ERA5 1991–2020 climatology, in the uncentered NCEP/EMC form so the numbers sit on the same scale as the standard charts. Higher is better; skill is generally spent by the time ACC falls through 0.6. How this works ▸

Skill matrix — last 30 days

Mean ACC per model × lead over the last 30 days of verifications — darker = sharper. The scorecard as a record, not a snapshot.

Regional skill — where each model wins, last 30 days

Mean ACC by region at Day 5 / 7 / 10 over the last 30 days — the spatial structure of model superiority. Row winner ringed.

Day

Autopsy

Run post-mortem

One run, end to end — its maps, its fork, its regime, the decisions that would have perfected it. Defaults to the newest run verified through Day 5; pick any recent init.

Init
Field

Anomaly correlation by lead

NH 20–80°N, cos-lat weighted · each model against its own analysis at the valid time (the operational-center convention; the common-truth board above instead grades everything against the one ECMWF analysis) · anomalies vs ERA5 1991–2020, uncentered ACC (NCEP/EMC form — the domain-mean anomaly is kept, as on the standard 5-day NH charts) · the 0.60 line marks where useful skill runs out

RMSE by lead

Root-mean-square height error (m) — the raw magnitude of the miss