Pith. sign in

Paper Citation Record · LEDGER

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation

As of 11 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2607.23493.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23493 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T20:58:37.097846Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6384f29a-f580-490b-987e-98c688113fcd · outbound

This paper cites The hateful memes challenge: Detecting hate speech in multimodal memes,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation The hateful memes challenge: Detecting hate speech in multimodal memes,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.051498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.051498Z digest=sha256:7f83984c573851decb163d5290741af1915b62dd88d2a6472b793bf1d9ccb93c

Observation 14d36f76-9da3-4d0e-a7f7-5414a5160d1a · outbound

This paper cites Knowmeme: A knowledge-enriched graph neural network solution to offensive meme detection,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Knowmeme: A knowledge-enriched graph neural network solution to offensive meme detection,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.059335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.059335Z digest=sha256:a31f4047b3faf15ca9a0e65af583a5ee8b21874cb953c6720dae2783a52c14df

Observation b25e1de4-76dc-444c-b2d4-4926f7497ece · outbound

This paper cites MOMENTA: A multimodal framework for detecting harmful memes and their targets,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation MOMENTA: A multimodal framework for detecting harmful memes and their targets,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.063002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.063002Z digest=sha256:8b780b8a5dce11eff648c031d11f4171b636c363dc057ec6fe86b70f0fd61ed2

Observation 08ccd657-b37e-4574-813a-9dfdda199423 · outbound

This paper cites DISARM: Detecting the Victims Targeted by Harmful Memes.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation DISARM: Detecting the Victims Targeted by Harmful Memes

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.066371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.066371Z digest=sha256:9d87b416f8637d41532485de8cd64a69feb4eb6b8fab7e23265c4155d4692d51

Observation 597214ae-edf3-428d-aac1-64c5177508d7 · outbound

This paper cites Prompting for multimodal hateful meme classification,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Prompting for multimodal hateful meme classification,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.070428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.070428Z digest=sha256:369dfc2070d8bd8db47a970e49bd1a685158d11a74216ff482d8f20445596e33

Observation 92a8e39d-f56b-487b-90c9-3cb2003d8040 · outbound

This paper cites Mapping memes to words for multimodal hateful meme classification,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Mapping memes to words for multimodal hateful meme classification,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.073581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.073581Z digest=sha256:86ce4eb3709fce7bb2b23f5102227ec33c41d0862204f2ff24ebfd69182aedea

Observation fe3b8366-e4a8-4ece-be94-cd056d4d4d0d · outbound

This paper cites Multimodal and Explainable Internet Meme Classification.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Multimodal and Explainable Internet Meme Classification

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.077724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.077724Z digest=sha256:096d5ca5985385f7a7877955e237b78dd54c06ca50c446055b43b20286453d23

Observation bb51594c-b494-4b3f-9827-a9de0a66ea79 · outbound

This paper cites Multimodal religiously hateful social media memes classification,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Multimodal religiously hateful social media memes classification,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.081190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.081190Z digest=sha256:7c923b64da2169c4bd93c93fcd7b289d662cc295359fa340908c725280ae95d6

Observation 3b705c5e-b529-4131-9358-98f55b61a4ce · outbound

This paper cites MUTE: A multimodal dataset for detecting hateful memes,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation MUTE: A multimodal dataset for detecting hateful memes,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.084729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.084729Z digest=sha256:5cfafa8a38985cc7e8946903307292641a786d4c487840e47c4733fbc4ff35b9

Observation 8e7abe3d-697b-4558-a865-990f2ca1a5c9 · outbound

This paper cites A multimodal framework to detect target aware aggression in memes,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation A multimodal framework to detect target aware aggression in memes,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.087823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.087823Z digest=sha256:395e6d707343a254ea389fb6856c0677efc73ba5bbd04744c2c32e3b5a7fead3

Observation e6f7dadc-167a-495e-a0f9-9a85e1f7b632 · outbound

This paper cites Deciphering hate: identifying hateful memes and their targets,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Deciphering hate: identifying hateful memes and their targets,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.091351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.091351Z digest=sha256:e5d7a56aa67b35a31f9465631837a01d86359527904f2e84498e7b39012da60f

Observation 29800500-4838-456f-b621-6c3b5bcb47f6 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Learning transferable visual models from natural language supervision,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.094556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.094556Z digest=sha256:e37eeddc6713e15a59957a0a398177c06daa7b0dee7056235877d69e4388240a

Observation d0585c94-15a3-4e32-8729-8e7033d03416 · outbound

This paper cites Unsupervised cross-lingual representation learning at scale,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Unsupervised cross-lingual representation learning at scale,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.097846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.097846Z digest=sha256:72d37279d21e97b7f8c69db24ba5e2afd439a1de3480f55247c9c500624e8f52

Pith citing papers

No inbound Pith citation observations are available.