Pith. sign in

REVIEW 1 cited by

BanglaAbuseMeme: A Dataset for Bengali Abusive Meme Classification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.11748 v1 pith:6NGQQB3P submitted 2023-10-18 cs.CV

classification cs.CV
keywords modelsmemesbengaliabusivedatasetbenchmarkbest-performingeffective
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The dramatic increase in the use of social media platforms for information sharing has also fueled a steep growth in online abuse. A simple yet effective way of abusing individuals or communities is by creating memes, which often integrate an image with a short piece of text layered on top of it. Such harmful elements are in rampant use and are a threat to online safety. Hence it is necessary to develop efficient models to detect and flag abusive memes. The problem becomes more challenging in a low-resource setting (e.g., Bengali memes, i.e., images with Bengali text embedded on it) because of the absence of benchmark datasets on which AI models could be trained. In this paper we bridge this gap by building a Bengali meme dataset. To setup an effective benchmark we implement several baseline models for classifying abusive memes using this dataset. We observe that multimodal models that use both textual and visual information outperform unimodal models. Our best-performing model achieves a macro F1 score of 70.51. Finally, we perform a qualitative error analysis of the misclassified memes of the best-performing text-based, image-based and multimodal models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. BTPD: A Multilingual Hand-curated Dataset of Bengali Transnational Political Discourse Across Online Communities

    cs.CL 2025-06 conditional novelty 6.0 of 10

    The paper presents BTPD, a new multilingual dataset of 2,235 hand-curated Bengali political posts from three online platforms, along with a descriptive topic overview.

Pith tools