Pith. sign in

Paper Citation Record · LEDGER

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation

As of 16 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2607.23493.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23493 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T20:58:37.097846Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6384f29a-f580-490b-987e-98c688113fcd · outbound

This paper cites The hateful memes challenge: Detecting hate speech in multimodal memes,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation The hateful memes challenge: Detecting hate speech in multimodal memes,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.051498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.051498Z digest=sha256:fd0688bdc3f9cefc13b13a2698a70300b71f6032aaf3881cab42e9aba4fc9fe8

Observation 14d36f76-9da3-4d0e-a7f7-5414a5160d1a · outbound

This paper cites Knowmeme: A knowledge-enriched graph neural network solution to offensive meme detection,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Knowmeme: A knowledge-enriched graph neural network solution to offensive meme detection,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.059335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.059335Z digest=sha256:25ebe73094eab8e6a801ebfa973a02646dfd5e267d9f6d059a88ad31c9550684

Observation b25e1de4-76dc-444c-b2d4-4926f7497ece · outbound

This paper cites MOMENTA: A multimodal framework for detecting harmful memes and their targets,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation MOMENTA: A multimodal framework for detecting harmful memes and their targets,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.063002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.063002Z digest=sha256:6c098953cd4a95679cccb74f4fcaf88ac5c6964aa8d5d90e2014897ffbc22224

Observation 08ccd657-b37e-4574-813a-9dfdda199423 · outbound

This paper cites DISARM: Detecting the Victims Targeted by Harmful Memes.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation DISARM: Detecting the Victims Targeted by Harmful Memes

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.066371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.066371Z digest=sha256:3b6c7c1cda64f405f7f3d331b0f658e38a1f8037c2805bfa8cb6921ab14d3a4e

Observation 597214ae-edf3-428d-aac1-64c5177508d7 · outbound

This paper cites Prompting for multimodal hateful meme classification,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Prompting for multimodal hateful meme classification,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.070428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.070428Z digest=sha256:81d78476bb17754ab21461d6a546d41a0f21c4040fe4df0e01fb45f11d202b9d

Observation 92a8e39d-f56b-487b-90c9-3cb2003d8040 · outbound

This paper cites Mapping memes to words for multimodal hateful meme classification,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Mapping memes to words for multimodal hateful meme classification,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.073581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.073581Z digest=sha256:2acdc0ff2f7faefc5af62e8b67d06610fbcd059fc538144d1af3ad4b334d489e

Observation fe3b8366-e4a8-4ece-be94-cd056d4d4d0d · outbound

This paper cites Multimodal and Explainable Internet Meme Classification.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Multimodal and Explainable Internet Meme Classification

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.077724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.077724Z digest=sha256:2d18d8e152b4f02908f92d4bceb03a144325186aa5b04b2f72c0dd92bf068e57

Observation bb51594c-b494-4b3f-9827-a9de0a66ea79 · outbound

This paper cites Multimodal religiously hateful social media memes classification,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Multimodal religiously hateful social media memes classification,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.081190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.081190Z digest=sha256:04ca509129bc21474e09e4390813e4867e54c37378dd8a928931b9a24f5be885

Observation 3b705c5e-b529-4131-9358-98f55b61a4ce · outbound

This paper cites MUTE: A multimodal dataset for detecting hateful memes,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation MUTE: A multimodal dataset for detecting hateful memes,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.084729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.084729Z digest=sha256:5287f2154dd04d8455317b0f5c8842f2f0dbc88c02a76ddd139cb4006693000c

Observation 8e7abe3d-697b-4558-a865-990f2ca1a5c9 · outbound

This paper cites A multimodal framework to detect target aware aggression in memes,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation A multimodal framework to detect target aware aggression in memes,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.087823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.087823Z digest=sha256:36331a445790ec4739934cf568ef34423979b211960fb1f6286cdbf953ecd4e7

Observation e6f7dadc-167a-495e-a0f9-9a85e1f7b632 · outbound

This paper cites Deciphering hate: identifying hateful memes and their targets,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Deciphering hate: identifying hateful memes and their targets,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.091351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.091351Z digest=sha256:133ec30e21f9d1a653c520e3659e8dbc19c8ee29364ee52f030fe44876ae181b

Observation 29800500-4838-456f-b621-6c3b5bcb47f6 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Learning transferable visual models from natural language supervision,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.094556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.094556Z digest=sha256:dfaf998b2a8ec6f98a9eeadb61a01760e004471b15458a0d1e1dafa3d91f7fa2

Observation d0585c94-15a3-4e32-8729-8e7033d03416 · outbound

This paper cites Unsupervised cross-lingual representation learning at scale,.

Token-Region Guided Cross-Attention Fusion for Multimodal Affect Interpretation Unsupervised cross-lingual representation learning at scale,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T20:58:37.097846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:58:37.097846Z digest=sha256:529cbfc12db1f461dd01cb92b3f122883e9db310a1af711f76708e7d1d7b3cc2

Pith citing papers

No inbound Pith citation observations are available.