Pith. sign in

Paper Citation Record · LEDGER

Supervised Multimodal Bitransformers for Classifying Images and Text

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:1909.02950.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1909.02950 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:23.736320Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:47:23.022501Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1cc77344-5872-4a58-9678-858a067c67ee · inbound

HuggingFace's Transformers: State-of-the-art Natural Language Processing cites this paper.

HuggingFace's Transformers: State-of-the-art Natural Language Processing Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 160

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:53:59.723692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T14:53:58.963468Z digest=sha256:8cf8ee4dfbebfc2169713b79c749e2d5b9255d2bc9c34749f7bc5b8661376e62

Observation 7328b916-8264-4ecb-aaf9-3e64ec6d9e21 · inbound

MemeReaCon: Probing Contextual Meme Understanding in Large Vision-Language Models cites this paper.

MemeReaCon: Probing Contextual Meme Understanding in Large Vision-Language Models Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:23.736320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:23.736320Z digest=sha256:dc145a1a058b2fdec76c689aae9bc94900cae58b61cfd634dff88baf320b28f1

Observation 4c4ec39b-7d84-4557-a1c6-48705d2ba659 · inbound

Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection cites this paper.

Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.720399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.720399Z digest=sha256:64317b1dd82403a8f8e818cca9575b140ddb1856a12782d1869334cfb89c8ced

Observation 50d67615-4a2c-486c-819d-73303de1f321 · inbound

Class Similarity-Based Multimodal Classification under Heterogeneous Category Sets cites this paper.

Class Similarity-Based Multimodal Classification under Heterogeneous Category Sets Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:37.092412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:37.092412Z digest=sha256:e7d4772a6de4cf9f0c9c1c65c098fa1c97d5390f5882e89d3f038895b7a19f20

Observation f2a86042-1463-404e-b5c5-2bf226c29f71 · inbound

AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on Harmfulness cites this paper.

AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on Harmfulness Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T20:51:25.874734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:51:25.874734Z digest=sha256:f2e1ea65dab18f579d9f49fa647a0179ce25722c09972152b27d1e7301682aef

Observation 96192f9b-b103-4de4-b12e-f9fbaba69c15 · inbound

MIND: A Multi-agent Framework for Zero-shot Harmful Meme Detection cites this paper.

MIND: A Multi-agent Framework for Zero-shot Harmful Meme Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:56:56.821826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:56:56.821826Z digest=sha256:58b58896aba4a2b34bb4f6570f695be7339a26367b6fba7504647a959f5e2ca1

Observation ac88227b-92e0-408c-b950-fae682fc0372 · inbound

EPIC: Efficient Prompt Interaction for Text-Image Classification cites this paper.

EPIC: Efficient Prompt Interaction for Text-Image Classification Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:46:00.157567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:46:00.157567Z digest=sha256:a97b1b89499f46143c130ec2561ed446a3d24f78da2eb4cd272ff4c7ef2f8e0b

Observation bb18bc7e-5361-47a0-84f3-51049671c499 · inbound

Applying multimodal learning to Classify transient Detections Early (AppleCiDEr) I: Data set, methods, and infrastructure cites this paper.

Applying multimodal learning to Classify transient Detections Early (AppleCiDEr) I: Data set, methods, and infrastructure Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T15:23:47.353424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:23:47.353424Z digest=sha256:67adb62cace90d5ce1d3cbcc736424b8ee61fd40b16cba1d176d586e41d39138

Observation b05b5ef7-ff1f-4302-9482-2c9f69152483 · inbound

Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval cites this paper.

Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:42:30.135543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T08:41:27.958069Z digest=sha256:d0590c55d49bbebfd222f45ca31675760dc4478b352193dce67172c367149c1e

Observation 7c352034-5ba7-422f-bc86-ab6d2e51163d · inbound

Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers cites this paper.

Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:31:00.451910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:28:55.932756Z digest=sha256:582e55f5207536a06e535a43bbaef30a34a62de11deb326986743253f850b303

Observation 73cbfa37-0a9a-4f64-9e0d-89a60603d956 · inbound

LongMoE: Longitudinal Multimodal Learning via Trajectory-Aware Mixture-of-Experts cites this paper.

LongMoE: Longitudinal Multimodal Learning via Trajectory-Aware Mixture-of-Experts Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.025038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T20:09:00.056162Z digest=sha256:ccc89923bedc88ab162147474c83fd388ebe78c91f091b32acba37162729d586

Observation b4ebeed2-5d38-403f-a6ef-35b55c6bcb1c · inbound

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection cites this paper.

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T06:07:51.123846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:07:51.123846Z digest=sha256:af6a153ea863660e0a17292ace503974dfdf2a1b6b605d4a9da4667892c3d500