Pith. sign in

REVIEW 3 cited by

The Unappreciated Role of Intent in Algorithmic Moderation of Social Media Content

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.11030 v1 pith:6CXZKSZO submitted 2024-05-17 cs.CL

The Unappreciated Role of Intent in Algorithmic Moderation of Social Media Content

classification cs.CL
keywords contentintentmoderationabusedetectionmodelscapturemedia
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

As social media has become a predominant mode of communication globally, the rise of abusive content threatens to undermine civil discourse. Recognizing the critical nature of this issue, a significant body of research has been dedicated to developing language models that can detect various types of online abuse, e.g., hate speech, cyberbullying. However, there exists a notable disconnect between platform policies, which often consider the author's intention as a criterion for content moderation, and the current capabilities of detection models, which typically lack efforts to capture intent. This paper examines the role of intent in content moderation systems. We review state of the art detection models and benchmark training datasets for online abuse to assess their awareness and ability to capture intent. We propose strategic changes to the design and development of automated detection and moderation systems to improve alignment with ethical and policy conceptualizations of abuse.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Safeguards Based on Copyable Context Cannot Provide Reliable Safety for LLMs

    cs.CR 2026-07 conditional novelty 7.0

    With copyable pre-release evidence, any dual-use release rule that keeps legitimate utility q must leave worst-case attacker assistance at least Γ(q)>0, so useful capability, reliable safety, and open access cannot coexist.

  2. Silence and Noise: Self-censorship and Opinion Expression on Social Media

    cs.SI 2026-04 unverdicted novelty 4.0

    Self-censorship on social media rises with larger audiences, lower posting frequency, and lower perceived support, causing users to align expressed views with perceived group norms.

  3. Silence and Noise: Self-censorship and Opinion Expression on Social Media

    cs.SI 2026-04 conditional novelty 4.0

    Self-censorship on social media rises with larger audiences, lower posting frequency, and lower perceived support, and speakers often align expressed views with perceived group norms.