Based on what you’ve seen so far, what are the main challenges with forecasting evaluation?
Based on what you’ve seen so far, what are the main challenges with forecasting evaluation? (my list)
Caution
When models have made forecasts for different subsets of tasks, it can make comparison of models complicated. For example, what if one model never forecasts for the harder prediction tasks?
Relative Skill Scores are a possible solution

“Comparing the accuracy of forecasting applications is difficult because forecasting methods, forecast outcomes, and reported validation metrics varied widely.”
– Chretien et al., PLOS ONE, 2014

Launched April 2020 by the Reich Lab in collaboration with CDC. Goals were to:
You will learn by doing in this session.
From a practical standpoint, standards facilitate
In 2024, CDC rebooted the hub, using hubverse standards.
These are the data that you will use in this session.
Use COVID-19 Forecast Hub forecasts to …
Forecast evaluation