SyncAI.news, a Varaisys broadcasting
WildfireSpreadBench: The Metric Decides the Model in Wildfire Spread Prediction
AG

Arin Gopakumar, Marco Pannozzo

· 1 min read

ResearcharXiv cs.LG

WildfireSpreadBench: The Metric Decides the Model in Wildfire Spread Prediction

arXiv:2609.22191v1 Announce Type: new Abstract: Machine learning is being increasingly used to predict where active wildfires will burn the following day, helping inform evacuation boundaries and containment lines. Most models are evaluated using Average Precision (AP), which summarizes performance across all decision thresholds, although acting on a forecast requires choosing one. We benchmarked five discriminative architectures and one generative model on WildfireSpreadTS using a shared evaluation pipeline and two input configurations. We found that model rankings varied depending on whether performance was measured by AP or by threshold-dependent metrics like F1 and IoU. The highest-AP model flagged 4 to 5 times the area that burned and ranked fifth of six on F1 and IoU, and the most recall-heavy model flagged 16 to 23 times. Models with more usable predictions had AP scores 24 to 37 lower. Across architectures, we identified three distinct prediction profiles: over-predicting, balanced, and under-predicting, which AP alone could not distinguish. Expanding the input from 7 to 23 channels changed AP by 0.03 on average, against a 0.21 to 0.24 spread across architectures. These results show AP alone can favor models whose predictions are poorly suited for operational wildfire forecasting.

Original source

This story was published by arXiv cs.LG and written by Arin Gopakumar, Marco Pannozzo. SyncAI.news shows a preview; the complete article is on the publisher's site.

Read the full story on arxiv.org

Similar News