Pith. sign in

Paper Citation Record · LEDGER

Supervised Multimodal Bitransformers for Classifying Images and Text

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:1909.02950.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1909.02950 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:34:47.070068Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:47:23.022501Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1cc77344-5872-4a58-9678-858a067c67ee · inbound

HuggingFace's Transformers: State-of-the-art Natural Language Processing cites this paper.

HuggingFace's Transformers: State-of-the-art Natural Language Processing Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 160

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:53:59.723692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-11T14:53:58.963468Z digest=sha256:4a58761a026f7f70fe852e439a3c0c64ce464ca9e79371e3e99c934141d59b7b

Observation 6fe500ab-a484-4dd5-b390-ff3486035b2d · inbound

Approximate Fiber Product: A Preliminary Algebraic-Geometric Perspective on Multimodal Embedding Alignment cites this paper.

Approximate Fiber Product: A Preliminary Algebraic-Geometric Perspective on Multimodal Embedding Alignment Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T05:34:47.070068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:34:47.070068Z digest=sha256:517ad4a4bb592afaf46b242d16ba3bd62610079b96fd0334cf765f861debdb4f

Observation 954efb43-2274-4a50-9fb5-0bf2a870adf3 · inbound

GAMED: Knowledge Adaptive Multi-Experts Decoupling for Multimodal Fake News Detection cites this paper.

GAMED: Knowledge Adaptive Multi-Experts Decoupling for Multimodal Fake News Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T17:41:29.492071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:41:29.492071Z digest=sha256:68bb1381598004d1bd68e058dd6441e284cc28a3958e51a5e8b218522b026a65

Observation 961357b2-5414-4f2e-9134-d69544604188 · inbound

MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data cites this paper.

MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T12:50:11.661278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:50:11.661278Z digest=sha256:f818ba914ab187f678ccd832a83f851cb957c2b2ad0b23e3b3b4b27c674e2414

Observation a9d303df-a774-4ec4-b9a8-145e834b8845 · inbound

Meme Trojan: Backdoor Attacks Against Hateful Meme Detection via Cross-Modal Triggers cites this paper.

Meme Trojan: Backdoor Attacks Against Hateful Meme Detection via Cross-Modal Triggers Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T11:25:10.506553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:25:10.506553Z digest=sha256:16d221524cc726238d8ce84f12b04197eb7aef964035a5059e985b42bc320d9d

Observation 7328b916-8264-4ecb-aaf9-3e64ec6d9e21 · inbound

MemeReaCon: Probing Contextual Meme Understanding in Large Vision-Language Models cites this paper.

MemeReaCon: Probing Contextual Meme Understanding in Large Vision-Language Models Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:23.736320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:51:23.736320Z digest=sha256:f47446ffc3f3152a1e2d550e75d05794193079ceccaa4eaa7286f5547d836359

Observation 4c4ec39b-7d84-4557-a1c6-48705d2ba659 · inbound

Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection cites this paper.

Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:39.720399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:39.720399Z digest=sha256:efd25b19e08504b068bf02e5e6b27462f74882debb3a66e992efb978be84e03a

Observation 50d67615-4a2c-486c-819d-73303de1f321 · inbound

Class Similarity-Based Multimodal Classification under Heterogeneous Category Sets cites this paper.

Class Similarity-Based Multimodal Classification under Heterogeneous Category Sets Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:37.092412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:37.092412Z digest=sha256:0d9a3a20ab066ca7bee226c3df7bf0755a620c6bce8e0ab248531d2b93a5d951

Observation f2a86042-1463-404e-b5c5-2bf226c29f71 · inbound

AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on Harmfulness cites this paper.

AdamMeme: Adaptively Probe the Reasoning Capacity of Multimodal Large Language Models on Harmfulness Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T20:51:25.874734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:51:25.874734Z digest=sha256:640029df04c93880d9d60273fffaa3561a80a7474ac6186869c2e3b5731710eb

Observation 96192f9b-b103-4de4-b12e-f9fbaba69c15 · inbound

MIND: A Multi-agent Framework for Zero-shot Harmful Meme Detection cites this paper.

MIND: A Multi-agent Framework for Zero-shot Harmful Meme Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T18:56:56.821826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:56:56.821826Z digest=sha256:afc943f9945b2c804d730f52ec7fca58cbf906059347fcf14d9df09c3c2e40f3

Observation ac88227b-92e0-408c-b950-fae682fc0372 · inbound

EPIC: Efficient Prompt Interaction for Text-Image Classification cites this paper.

EPIC: Efficient Prompt Interaction for Text-Image Classification Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:46:00.157567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:46:00.157567Z digest=sha256:acbf0e17680c7d285182c451f428ca354adba0c8915dd261522507e09c92f319

Observation bb18bc7e-5361-47a0-84f3-51049671c499 · inbound

Applying multimodal learning to Classify transient Detections Early (AppleCiDEr) I: Data set, methods, and infrastructure cites this paper.

Applying multimodal learning to Classify transient Detections Early (AppleCiDEr) I: Data set, methods, and infrastructure Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T15:23:47.353424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:23:47.353424Z digest=sha256:2539d28c4fd21b4388d015b81be6bc8e9e5f842d20d8e0e762d4559dc4c6cb70

Observation b05b5ef7-ff1f-4302-9482-2c9f69152483 · inbound

Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval cites this paper.

Fall into a Pit, Gain in a Wit: Cognitive-Guided Harmful Meme Detection via Misjudgment Risk Pattern Retrieval Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:42:30.135543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T08:41:27.958069Z digest=sha256:93cc742ded05f31c5c0d572e7fecad92057d6ec021ad4f195063e8253066c531

Observation 7c352034-5ba7-422f-bc86-ab6d2e51163d · inbound

Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers cites this paper.

Connecting online criminal behavior with machine learning: Using authorship attribution to analyze and link potential online traffickers Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:31:00.451910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T15:28:55.932756Z digest=sha256:32cfe7ca86baa24870ce65aa0456e01fcc1afeab782ed687e406a79a9caa91ac

Observation 73cbfa37-0a9a-4f64-9e0d-89a60603d956 · inbound

LongMoE: Longitudinal Multimodal Learning via Trajectory-Aware Mixture-of-Experts cites this paper.

LongMoE: Longitudinal Multimodal Learning via Trajectory-Aware Mixture-of-Experts Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.025038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T20:09:00.056162Z digest=sha256:7bb4b45eb43c6d22643bf56623f38ca523ebc9a45a842f3755497e93f81032b3

Observation b4ebeed2-5d38-403f-a6ef-35b55c6bcb1c · inbound

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection cites this paper.

EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection Supervised Multimodal Bitransformers for Classifying Images and Text

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T06:07:51.123846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:07:51.123846Z digest=sha256:61c1e0439c31dc8458b41aff2ba2ea95370544abab1c69c0e23f6cd1e05c0029