CrisisBench: Benchmarking Crisis-related Social Media Datasets for Humanitarian Information Processing

Authors

Firoj Alam,Hassan Sajjad,Muhammad Imran,Ferda Ofli

QCRI, Doha, Qatar,QCRI, Doha, Qatar,QCRI, Doha, Qatar,QCRI, Doha, Qatar

Proceedings:

Vol. 15 (2021): Fifteenth International AAAI Conference on Web and Social Media

Volume

Issue:

Vol. 15 (2021): Fifteenth International AAAI Conference on Web and Social Media

Track:

Dataset Papers

Downloads:

Download PDF

Abstract:

The time-critical analysis of social media streams is important for humanitarian organizations to plan rapid response during disasters. The crisis informatics research community has developed several techniques and systems to process and classify big crisis-related data posted on social media. However, due to the dispersed nature of the datasets used in the literature, it is not possible to compare the results and measure the progress made towards better models for crisis informatics. In this work, we attempt to bridge this gap by combining various existing crisis-related datasets. We consolidate eight annotated data sources and provide 166.1k and 141.5k tweets for informativeness and humanitarian classification tasks, respectively. The consolidation results in a larger dataset that affords the ability to train more sophisticated models. To that end, we provide binary and multiclass classification results using CNN, FastText, and transformer based models to address informativeness and humanitarian tasks, respectively. We make the dataset and scripts available at https://crisisnlp.qcri.org/crisis_datasets_benchmarks.html.

DOI:

10.1609/icwsm.v15i1.18115

ICWSM

Vol. 15 (2021): Fifteenth International AAAI Conference on Web and Social Media

Cookie	Duration	Description
cookielawinfo-checkbox-analytics	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Analytics".
cookielawinfo-checkbox-functional	11 months	The cookie is set by GDPR cookie consent to record the user consent for the cookies in the category "Functional".
cookielawinfo-checkbox-necessary	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookies is used to store the user consent for the cookies in the category "Necessary".
cookielawinfo-checkbox-others	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Other.
cookielawinfo-checkbox-performance	11 months	This cookie is set by GDPR Cookie Consent plugin. The cookie is used to store the user consent for the cookies in the category "Performance".
viewed_cookie_policy	11 months	The cookie is set by the GDPR Cookie Consent plugin and is used to store whether or not user has consented to the use of cookies. It does not store any personal data.