Pith. sign in

REVIEW 2 cited by

Multi3Hate: Multimodal, Multilingual, and Multicultural Hate Speech Detection with Vision-Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2411.03888 v2 pith:TDQ7R3VO submitted 2024-11-06 cs.CL

classification cs.CL
keywords hatemultimodalspeechculturaldatasetmodelsmulti3hatemultilingual
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Warning: this paper contains content that may be offensive or upsetting Hate speech moderation on global platforms poses unique challenges due to the multimodal and multilingual nature of content, along with the varying cultural perceptions. How well do current vision-language models (VLMs) navigate these nuances? To investigate this, we create the first multimodal and multilingual parallel hate speech dataset, annotated by a multicultural set of annotators, called Multi3Hate. It contains 300 parallel meme samples across 5 languages: English, German, Spanish, Hindi, and Mandarin. We demonstrate that cultural background significantly affects multimodal hate speech annotation in our dataset. The average pairwise agreement among countries is just 74%, significantly lower than that of randomly selected annotator groups. Our qualitative analysis indicates that the lowest pairwise label agreement-only 67% between the USA and India-can be attributed to cultural factors. We then conduct experiments with 5 large VLMs in a zero-shot setting, finding that these models align more closely with annotations from the US than with those from other cultures, even when the memes and prompts are presented in the dominant language of the other culture. Code and dataset are available at https://github.com/MinhDucBui/Multi3Hate.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. On Generalization across Measurement Systems: LLMs Entail More Test-Time Compute for Underrepresented Cultures

    cs.CL 2025-06 conditional novelty 6.0 of 10

    LLMs are less accurate when asked to report facts in non-default measurement systems, and chain-of-thought restores accuracy only at a 180-300 percent increase in test-time compute.

  2. Unified Game Moderation: Soft-Prompting and LLM-Assisted Label Transfer for Resource-Efficient Toxicity Detection

    cs.CL 2025-06 conditional novelty 4.0 of 10

    A single BERT-scale model with a game-context token and LLM-assisted label transfer achieves toxicity detection comparable to per-game models while extending to seven languages.

Pith tools