Pith. sign in

REVIEW 2 cited by

Hierarchical Multi-Label Classification of Online Vaccine Concerns

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.01783 v1 pith:OSMEF76I submitted 2024-02-01 cs.CL cs.AIcs.LG

classification cs.CLcs.AIcs.LG
keywords concernsvaccineonlinedifferentexploreinformpromptingstrategies
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Vaccine concerns are an ever-evolving target, and can shift quickly as seen during the COVID-19 pandemic. Identifying longitudinal trends in vaccine concerns and misinformation might inform the healthcare space by helping public health efforts strategically allocate resources or information campaigns. We explore the task of detecting vaccine concerns in online discourse using large language models (LLMs) in a zero-shot setting without the need for expensive training datasets. Since real-time monitoring of online sources requires large-scale inference, we explore cost-accuracy trade-offs of different prompting strategies and offer concrete takeaways that may inform choices in system designs for current applications. An analysis of different prompting strategies reveals that classifying the concerns over multiple passes through the LLM, each consisting a boolean question whether the text mentions a vaccine concern or not, works the best. Our results indicate that GPT-4 can strongly outperform crowdworker accuracy when compared to ground truth annotations provided by experts on the recently introduced VaxConcerns dataset, achieving an overall F1 score of 78.7%.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Can Large Language Models Serve as Effective Classifiers for Hierarchical Multi-Label Classification of Scientific Documents at Industrial Scale?

    cs.AI 2024-12 reject novelty 5.0 of 10

    A retrieval plus zero-shot LLM pipeline is reported to give 94.3% SME-approval accuracy on SSRN hierarchical multi-label classification, versus 61.5% for fine-tuned SPECTER2, with no retraining.

  2. A Platform for Investigating Public Health Content with Efficient Concern Classification

    cs.CL 2025-06 conditional novelty 4.0 of 10

    The paper introduces ConcernScope, a teacher-student platform where GPT-4 labels training data and a BERT model classifies texts into VaxConcerns categories, with a pilot trend analysis on 186,000 passages.

Pith tools