DisastIR: A Comprehensive Information Retrieval Benchmark for Disaster Management

Kai Yin; Xiangjue Dong; Chengkai Liu; Lipai Huang; Yiming Xiao; Zhewei Liu; Ali Mostafavi; James Caverlee

doi:10.18653/v1/2025.findings-emnlp.97

DisastIR: A Comprehensive Information Retrieval Benchmark for Disaster Management

Kai Yin, Xiangjue Dong, Chengkai Liu, Lipai Huang, Yiming Xiao, Zhewei Liu, Ali Mostafavi, James Caverlee

Abstract

Effective disaster management requires timely access to accurate and contextually relevant information. Existing Information Retrieval (IR) benchmarks, however, focus primarily on general or specialized domains, such as medicine or finance, neglecting the unique linguistic complexity and diverse information needs encountered in disaster management scenarios. To bridge this gap, we introduce DisastIR, the first comprehensive IR evaluation benchmark specifically tailored for disaster management. DisastIR comprises 9,600 diverse user queries and more than 1.3 million labeled query-passage pairs, covering 48 distinct retrieval tasks derived from six search intents and eight general disaster categories that include 301 specific event types. Our evaluations of 30 state-of-the-art retrieval models demonstrate significant performance variances across tasks, with no single model excelling universally. Furthermore, comparative analyses reveal significant performance gaps between general-domain and disaster management-specific tasks, highlighting the necessity of disaster management-specific benchmarks for guiding IR model selection to support effective decision-making in disaster management scenarios. All source codes and DisastIR are available at https://github.com/KaiYin97/Disaster_IR.

Anthology ID:: 2025.findings-emnlp.97
Volume:: Findings of the Association for Computational Linguistics: EMNLP 2025
Month:: November
Year:: 2025
Address:: Suzhou, China
Editors:: Christos Christodoulopoulos, Tanmoy Chakraborty, Carolyn Rose, Violet Peng
Venue:: Findings
SIG:
Publisher:: Association for Computational Linguistics
Note:
Pages:: 1836–1867
Language:
URL:: https://aclanthology.org/2025.findings-emnlp.97/
DOI:: 10.18653/v1/2025.findings-emnlp.97
Bibkey:
Cite (ACL):: Kai Yin, Xiangjue Dong, Chengkai Liu, Lipai Huang, Yiming Xiao, Zhewei Liu, Ali Mostafavi, and James Caverlee. 2025. DisastIR: A Comprehensive Information Retrieval Benchmark for Disaster Management. In Findings of the Association for Computational Linguistics: EMNLP 2025, pages 1836–1867, Suzhou, China. Association for Computational Linguistics.
Cite (Informal):: DisastIR: A Comprehensive Information Retrieval Benchmark for Disaster Management (Yin et al., Findings 2025)
Copy Citation:
PDF:: https://aclanthology.org/2025.findings-emnlp.97.pdf
Checklist:: 2025.findings-emnlp.97.checklist.pdf

PDF Cite Search Checklist Fix data