REVIEW 2 cited by
Large Scale Crowdsourcing and Characterization of Twitter Abusive Behavior
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In recent years, offensive, abusive and hateful language, sexism, racism and other types of aggressive and cyberbullying behavior have been manifesting with increased frequency, and in many online social media platforms. In fact, past scientific work focused on studying these forms in popular media, such as Facebook and Twitter. Building on such work, we present an 8-month study of the various forms of abusive behavior on Twitter, in a holistic fashion. Departing from past work, we examine a wide variety of labeling schemes, which cover different forms of abusive behavior, at the same time. We propose an incremental and iterative methodology, that utilizes the power of crowdsourcing to annotate a large scale collection of tweets with a set of abuse-related labels. In fact, by applying our methodology including statistical analysis for label merging or elimination, we identify a reduced but robust set of labels. Finally, we offer a first overview and findings of our collected and annotated dataset of 100 thousand tweets, which we make publicly available for further scientific exploration.
Forward citations
Cited by 2 Pith papers
-
Patterns and Purposes: A Cross-Journal Analysis of AI Tool Usage in Academic Writing
Across 168 AI-use declarations from 2024 Elsevier articles, ChatGPT is the dominant tool and readability improvement is the top stated purpose, with reported differences between author groups.
-
Evaluating Simple Debiasing Techniques in RoBERTa-based Hate Speech Detection Models
Simple debiasing reduces dialect disparity in RoBERTa hate speech models only when training data is balanced across dialect subgroups.
Discussion (0). Continue with ORCID to comment.