fix/corrected-benchmark-and-lag,
return levels fitted strictly inside 1999–2023).
The two configurations
Each window's voting units are listed under its rule — a gauge paired with the product read at it, with its rank correlation against the reference gauge and its best lag in days. These are what survived selection; the count of candidates considered is shown beside each rule.
One asymmetry worth naming. Because the relative floor runs within each product for the one-model setup, that setup can reach units the mixed setup rejects as unfit. It bites in exactly one window: in Shabelle Deyr, Google's three gauges sit at ρ 0.62–0.64 against GloFAS at Belet Weyne's 0.83, so the mixed pool cuts all three (3 of 6 candidates) while the one-model setup is offered Google as a complete three-gauge option. It declined it — both setups adopt the same three GloFAS units there, so nothing on this page turns on the choice. The absolute ρ ≥ 0.50 floor applies either way, so the per-product floor can admit a mediocre unit but never a poor one.
| loading… |
1. Historical trigger record
One row per year, one block of columns per basin-season. The number is how many voting units were over their own thresholds at once; a cell is shaded when it reaches the threshold printed in its own column header, which is exactly when that window activates. Thresholds differ by column, so the same number can be shaded in one window and left plain in another. After the two trigger columns each block carries what actually happened: the gauge class, the CERF flood allocation and the EM-DAT people affected attributed to that basin-season, and whether the year is on the target list.
| loading… |
2. Summary statistics
Scored two ways. Against the gauge benchmark — the years in which two or more of a river's gauges reached a 1-in-5 level in that season — and against a constructed target list of years the framework arguably should have paid out in: . Both are computed per basin-season and then for the envelope. Lead time is the mean number of days between the first forecast alert at leads 1–6 and the reference gauge crossing its own 1-in-3 level; negative means the trigger fired after the river was already up.
| loading… |
The target years
| loading… |
CERF allocations naming neither river are counted under both basins. EM-DAT events are attributed by the river named in the location text and split to seasons by start month, so a nationwide event appears under both basins. This list is a proxy for “years the framework should have paid out”, not a validated impact record: it inherits EM-DAT's entry-criteria bias and the fact that a CERF allocation reflects an appeal as much as a flood.