Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:55:28.327964Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2505.00150.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:55:28.327964Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 27a3026c-8019-4ea1-8265-bac079e7f749 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Flamingo: a visual language model for few-shot learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 715f2511-14f1-45ad-91ea-94d41e0d9a33 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab681ece-2041-4fd9-ac45-237358c941fb · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Prompting for Multimodal Hateful Meme Classification
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dae2445-6fe6-4f36-ac7e-5c76545edf3d · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Modularized networks for few-shot hateful meme detection
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8cc59cd8-b6d4-4ccb-979a-ad3a279ab324 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17e15a67-b841-44b9-978a-f1a048b1bb29 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58749952-d8c6-4df9-8cc9-7cb8672a5cda · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Deep residual learning for image recognition
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d29ebf04-445a-4509-8709-b2dfbb4d99bb · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Recent advances in online hate speech moderation: Multimodality and the role of large models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b0a87666-f99c-40cb-ab81-2bd4a9bac7fc · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Openclip, 2021
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5d1c160-87dd-4d46-b1b1-721aacc15379 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Capalign: Improving cross modal alignment via informative captioning for harmful meme detection
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4784a6f9-9cad-4d84-9cc5-9d12b36057ed · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Supervised Multimodal Bitransformers for Classifying Images and Text
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90167aa6-9778-4423-819e-70e3b590095b · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models The hateful memes challenge: Detecting hate speech in multimodal memes
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6da361d-0501-424c-a64e-90b3f5094bed · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models The hateful memes challenge: Competition report
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 83116132-e221-472d-99c3-835d39beed2a · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Why is it hate speech? masked ratio- nale prediction for explainable hate speech detection
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cfe61be9-246c-491d-ba36-3ce256c73838 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Segment anything
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2ca12f7-87b6-4278-a7c3-a1028ff340de · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a673173-6d41-46a3-92b8-bc454f84d70f · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models VisualBERT: A Simple and Performant Baseline for Vision and Language
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27cd51f3-2628-4dbd-b8cf-95fce6661f8f · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Towards explainable harmful meme detection through multimodal debate between large language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cb0c1590-2025-4a6d-91dd-c5cd6381f695 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models A Multimodal Framework for the Detection of Hateful Memes
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12607ca3-8f04-45db-b528-c90674b99931 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Visual Instruction Tuning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92183d9d-f78a-4287-ae26-597e4739640c · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d433d1-594f-41b7-90f4-22b6da9fb43c · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d60a65-1c98-4d21-8c04-49174206c2b1 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Hatexplain: A benchmark dataset for explainable hate speech detection
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 32d06113-ec91-4234-ae51-93dbe547cd37 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models ETHOS: an Online Hate Speech Detection Dataset
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e834cf9-0d87-4b47-b0bf-850ddca0a227 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Vilio: State-of-the-art Visio-Linguistic Models applied to Hateful Memes
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec4bab69-10c7-4529-92a0-14085d108fb3 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models GPT-4 Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f8bd0e9-2177-4d03-b8dd-4e67ba36e8e1 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Unsafe diffusion: On the generation of unsafe images and hateful memes from text-to-image models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da24017-0a3d-41c0-b2ae-fa22a565caab · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Learning transferable visual models from natural language supervision
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d9b3b92-8ba6-4e34-8337-a62d3532c2dc · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c781a117-7ef3-44fe-929b-d80fef9691a6 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Detecting Hateful Memes Using a Multimodal Deep Ensemble
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eda2e25-6c3e-42d7-8b47-906f7f2afa7b · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Detecting formal thought disorder by deep contextualized word representations
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ecba3b8f-09f6-407b-979a-efc115661f23 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 709d1768-4ad3-4869-8d48-0abb639ea700 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Large Language Models Encode Clinical Knowledge
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e31eb3b-a8a7-4db6-b683-1e64fc33f7f8 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Resolution-robust large mask inpainting with fourier convolutions
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e860ffdb-c39d-4dfa-878e-80d67f035e85 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe78932e-7f8f-4f0b-849d-807cb4ac6294 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b00b850-8e44-4985-8f20-bf7c41cd78e9 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models On large visual language models for medical imaging analysis: An empirical study
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 68a69b57-5211-4cec-aab3-d4b763283f99 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Detecting Hate Speech in Memes Using Multimodal Deep Learning Approaches: Prize-winning solution to Hateful Memes Challenge
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 176e6f5b-20fb-42bf-9ac9-65154d5a5c81 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Memecraft: Contextual and stance-driven multimodal meme generation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d48cd019-2e12-462d-9843-8ecb8b589f24 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models NExT-GPT: Any-to-Any Multimodal LLM
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9ce84d-1427-4cd6-bc88-1e6c9aee62de · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Aggregated residual transformations for deep neural networks
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c3ae4b9-e547-49c0-8cfa-c47ea4533354 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Coded hate speech detection via contextual information
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5bb76d33-83f3-434a-a6c0-ce0a5f822d11 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models An empirical study of gpt-3 for few-shot knowledge-based vqa
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 041eb16d-2f38-4db5-a17d-8deb554b4ecc · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Inpaint Anything: Segment Anything Meets Image Inpainting
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa5ea7f7-4550-4bbf-a441-ac8e85e1a750 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42b50c13-34e7-4099-9e99-690d94fb0514 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models OPT: Open Pre-trained Transformer Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e316939-a3d1-44fd-ab84-4eeaa6169eb7 · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models Enhance Multimodal Transformer With External Label And In-Domain Pretrain: Hateful Meme Challenge Winning Solution
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1570484-a50e-4715-a304-3ca31926a19e · outbound
Detecting and Mitigating Hateful Content in Multimodal Memes with Vision-Language Models C.2 Multimodal: unimodal pretraining Multimodal models from unimodal pretraining typically combine the output or the features of vision and linguistic models
Reference 768
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
No inbound Pith citation observations are available.