Pith. sign in

REVIEW 1 cited by

Zero-Shot Classification of Crisis Tweets Using Instruction-Finetuned Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.00182 v1 pith:3CZIGQEM submitted 2024-09-30 cs.CL cs.AI

classification cs.CLcs.AI
keywords classificationmodelscrisisdatasethumanitarianpostsbetterevaluated
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Social media posts are frequently identified as a valuable source of open-source intelligence for disaster response, and pre-LLM NLP techniques have been evaluated on datasets of crisis tweets. We assess three commercial large language models (OpenAI GPT-4o, Gemini 1.5-flash-001 and Anthropic Claude-3-5 Sonnet) capabilities in zero-shot classification of short social media posts. In one prompt, the models are asked to perform two classification tasks: 1) identify if the post is informative in a humanitarian context; and 2) rank and provide probabilities for the post in relation to 16 possible humanitarian classes. The posts being classified are from the consolidated crisis tweet dataset, CrisisBench. Results are evaluated using macro, weighted, and binary F1-scores. The informative classification task, generally performed better without extra information, while for the humanitarian label classification providing the event that occurred during which the tweet was mined, resulted in better performance. Further, we found that the models have significantly varying performance by dataset, which raises questions about dataset quality.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Harnessing Large Language Models for Disaster Management: A Survey

    cs.CL 2025-01 conditional novelty 4.0 of 10

    A review and taxonomy of large language model applications for natural disaster management, with a public dataset catalog and a list of research challenges.

Pith tools