Zero-Shot Classification of Crisis Tweets Using Instruction-Finetuned Large Language Models

Fuente: arXiv
Guardado en:
Detalles Bibliográficos
Autores principales: McDaniel, Emma, Scheele, Samuel, Liu, Jeff
Formato: Preprint
Publicado: 2024
Materias:
Acceso en línea:
Etiquetas: Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
_version_ 1866911125927362560
author McDaniel, Emma
Scheele, Samuel
Liu, Jeff
author_facet McDaniel, Emma
Scheele, Samuel
Liu, Jeff
contents Social media posts are frequently identified as a valuable source of open-source intelligence for disaster response, and pre-LLM NLP techniques have been evaluated on datasets of crisis tweets. We assess three commercial large language models (OpenAI GPT-4o, Gemini 1.5-flash-001 and Anthropic Claude-3-5 Sonnet) capabilities in zero-shot classification of short social media posts. In one prompt, the models are asked to perform two classification tasks: 1) identify if the post is informative in a humanitarian context; and 2) rank and provide probabilities for the post in relation to 16 possible humanitarian classes. The posts being classified are from the consolidated crisis tweet dataset, CrisisBench. Results are evaluated using macro, weighted, and binary F1-scores. The informative classification task, generally performed better without extra information, while for the humanitarian label classification providing the event that occurred during which the tweet was mined, resulted in better performance. Further, we found that the models have significantly varying performance by dataset, which raises questions about dataset quality.
format Preprint
id arxiv_https___arxiv_org_abs_2410_00182
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Zero-Shot Classification of Crisis Tweets Using Instruction-Finetuned Large Language Models
McDaniel, Emma
Scheele, Samuel
Liu, Jeff
Computation and Language
Artificial Intelligence
Social media posts are frequently identified as a valuable source of open-source intelligence for disaster response, and pre-LLM NLP techniques have been evaluated on datasets of crisis tweets. We assess three commercial large language models (OpenAI GPT-4o, Gemini 1.5-flash-001 and Anthropic Claude-3-5 Sonnet) capabilities in zero-shot classification of short social media posts. In one prompt, the models are asked to perform two classification tasks: 1) identify if the post is informative in a humanitarian context; and 2) rank and provide probabilities for the post in relation to 16 possible humanitarian classes. The posts being classified are from the consolidated crisis tweet dataset, CrisisBench. Results are evaluated using macro, weighted, and binary F1-scores. The informative classification task, generally performed better without extra information, while for the humanitarian label classification providing the event that occurred during which the tweet was mined, resulted in better performance. Further, we found that the models have significantly varying performance by dataset, which raises questions about dataset quality.
title Zero-Shot Classification of Crisis Tweets Using Instruction-Finetuned Large Language Models
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2410.00182