Most weather apps don't tell you how often they're wrong. We do. Every three hours we snapshot what each of our five forecasts said. Every hour we fetch actual readings from UK Met Office stations. Then we score the forecasts against what actually happened. The numbers below are the result.
Four numbers per source. Each is colour-coded against what "good" actually means for that measure — they're deliberately different bars, because some things are much harder to forecast than others:
"Our blend" is the forecast the app actually shows you. It weighs the five sources by their proven track record at each measure and each time horizon, and re-weighs itself automatically as the scores evolve — so if one source goes off the boil, the blend leans away from it without anyone touching anything. The tick panel at the top tells you at a glance whether the blend is currently beating the individual forecasts.
We pull observations from Open-Meteo's Historical Weather Archive, which blends ground-station readings with ECMWF's ERA5 reanalysis to give an hourly observation record for any point in the UK. We monitor a set of locations spread right across the UK — from the south coast to Edinburgh and Belfast — so the scores reflect performance nationally, not just one corner.
We don't cherry-pick. We don't drop bad runs. We don't weight scores by what makes any one source look better. By default we show every day we've scored since tracking began; you can narrow the window with a ?days=N URL parameter.