REVIEW 4 cited by
Benchmarking Zero-shot Text Classification: Datasets, Evaluation and Entailment Approach
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Zero-shot text classification (0Shot-TC) is a challenging NLU problem to which little attention has been paid by the research community. 0Shot-TC aims to associate an appropriate label with a piece of text, irrespective of the text domain and the aspect (e.g., topic, emotion, event, etc.) described by the label. And there are only a few articles studying 0Shot-TC, all focusing only on topical categorization which, we argue, is just the tip of the iceberg in 0Shot-TC. In addition, the chaotic experiments in literature make no uniform comparison, which blurs the progress. This work benchmarks the 0Shot-TC problem by providing unified datasets, standardized evaluations, and state-of-the-art baselines. Our contributions include: i) The datasets we provide facilitate studying 0Shot-TC relative to conceptually different and diverse aspects: the ``topic'' aspect includes ``sports'' and ``politics'' as labels; the ``emotion'' aspect includes ``joy'' and ``anger''; the ``situation'' aspect includes ``medical assistance'' and ``water shortage''. ii) We extend the existing evaluation setup (label-partially-unseen) -- given a dataset, train on some labels, test on all labels -- to include a more challenging yet realistic evaluation label-fully-unseen 0Shot-TC (Chang et al., 2008), aiming at classifying text snippets without seeing task specific training data at all. iii) We unify the 0Shot-TC of diverse aspects within a textual entailment formulation and study it this way. Code & Data: https://github.com/yinwenpeng/BenchmarkingZeroShot
Forward citations
Cited by 4 Pith papers
-
Social Contagion in COVID-19 Discussions within the Belgian Reddit Community: A Statistical and Modeling Study
In r/Belgium, COVID-19 topics were seeded by external events, not by prior posts, but comment sentiment was contagious, and a two-layer bounded confidence model best captured that asymmetry.
-
A Modular Unsupervised Framework for Attribute Recognition from Unstructured Text
POSID combines regex, Word2Vec, WordNet, and SBERT zero-shot classification with POS-tag heuristics to extract person attributes from incident reports, reaching 0.90 F1 for clothes attribute-value pairs on the new Inc...
-
Multimodal Information Retrieval for Open World with Edit Distance Weak Supervision
FemmIR uses graph-edit-distance weak supervision over extracted object properties to rank multimodal retrieval results without any similarity labels or fine-tuning.
-
Beyond Traditional Algorithms: Leveraging LLMs for Accurate Cross-Border Entity Identification
A 65-case comparison claims commercial chatbot LLMs are the most accurate for Portuguese entity matching, but the reported false-positive rates contradict the claim.
Discussion (0). Sign in to comment.