Pith. sign in

Paper Citation Record · LEDGER

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

As of 22 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 2 inbound Pith citation observations for arXiv:2509.17446.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.17446 v4

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:46.996181Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:51:46.891087Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-08T00:49:09.573510Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f528d143-3c65-4591-bbea-a709f30617cc · outbound

This paper cites MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.891087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.891087Z digest=sha256:94d6bad2d0dc2ef3e9da269dad6e2169bcc466b692923a7c91bf2528170caa8c

Observation 5567a18e-07cc-4aba-ace6-2e1514b45ff2 · outbound

This paper cites Model Overview As illustrated in Fig.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Model Overview As illustrated in Fig

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.546471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.896312Z digest=sha256:5f82d11c1334b7aaac5e387a4bb63bd8cf9077c50f5f7ad87fe24312858d5880

Observation 90a68bb4-30f6-4c65-8d03-33eea146f003 · outbound

This paper cites an unresolved cited work.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Unresolved cited work

Reference 3

Resolution
verified exact
raw_fallback, observed 2026-08-15T15:51:47.273979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.901222Z digest=sha256:703ba2e076b3d27369db0c28eb5d7059638f9066bc2ec891a90f3fc532ac9a75

Observation b0f42847-ed54-4637-a397-574b2270c050 · outbound

This paper cites an unresolved cited work.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:47.532950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.906048Z digest=sha256:e372d64bfd3ef175293377450650bd6dd910c573169ff3b08b21dd2350312b42

Observation 2d4d5bcf-f8a5-4021-bbde-ae522dfce194 · outbound

This paper cites an unresolved cited work.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-15T15:51:47.519408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.910514Z digest=sha256:7047b8a32f355757feac9fd554ae17935bbeb866810c10ffcc749196f5d4c2f6

Observation e48c8a52-375c-4de8-9d13-b36269182ee9 · outbound

This paper cites Deep Learning Approaches for Multimodal Intent Recognition: A Survey.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Deep Learning Approaches for Multimodal Intent Recognition: A Survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.914985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.914985Z digest=sha256:df21b43ee881407151ef6db8e765ddc8f37205949e3bf0a056f0d41ecf2ad550

Observation bc6046f5-a54f-421d-9d01-f2a1a218b18c · outbound

This paper cites Temporal working memory: Query- guided segment refinement for enhanced multimodal understanding,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Temporal working memory: Query- guided segment refinement for enhanced multimodal understanding,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.506674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.919561Z digest=sha256:b6ee7693dc6964a6d9bff306696c7b105a3a40b931dc7da57125ed90f821c6b9

Observation 2b22f4fc-0697-4edd-a33c-aa678ca08fa9 · outbound

This paper cites To- wards visual-prompt temporal answer grounding in in- structional video,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion To- wards visual-prompt temporal answer grounding in in- structional video,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.494272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.923324Z digest=sha256:5f8eacf2a6670d13b8764d6de5966c30a5d7611e2b5c549660c095e7f7dfec57

Observation f4672e7e-dde4-42b1-9222-ebcebf606a84 · outbound

This paper cites Visual document understand- ing and question answering: A multi-agent collabora- tion framework with test-time scaling,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Visual document understand- ing and question answering: A multi-agent collabora- tion framework with test-time scaling,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.926883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.926883Z digest=sha256:40e30d31310e47aa4c86cc47ac8c49856cb590035a229ab27eae805b19a6b977

Observation de325dd4-b4e6-40af-b9dd-ddb49be69200 · outbound

This paper cites Multimodal transformer for un- aligned multimodal language sequences,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Multimodal transformer for un- aligned multimodal language sequences,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.480557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.930606Z digest=sha256:caa434a2b78c6088f9c2f188cd020f64933b56ab0104e9086ec517f3f184e7e5

Observation 5b2d2083-93f1-4545-9bcc-b0812502239d · outbound

This paper cites Adaptive mul- timodal fusion: Dynamic attention allocation for intent recognition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Adaptive mul- timodal fusion: Dynamic attention allocation for intent recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.467553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.934791Z digest=sha256:66acca6d36479c4c3804c7cf472b9bc656027c6f27f4ad08a862170bcb43368a

Observation 6000014f-e138-4efc-b697-2d60ff710782 · outbound

This paper cites Multimodal transformer with multi-scale alignment for multimodal sentiment analysis,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Multimodal transformer with multi-scale alignment for multimodal sentiment analysis,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.454648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.938626Z digest=sha256:629ef2896bf0eba42eb35446996492d601594b2c33487df0faefd58149b36b07

Observation 7da4ab5b-a2fa-4629-8a4b-a046ceb2c561 · outbound

This paper cites Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Improving Multimodal Fusion with Hierarchical Mutual Information Maximization for Multimodal Sentiment Analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.942355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.942355Z digest=sha256:a98df8e4bab5ffd454796233b72a7cbd2b33fe166f97c00ae8b11e6b0dc654de

Observation 0ddd00f2-d4fa-4101-9bce-1ed40ecc9cc9 · outbound

This paper cites Dynamic multimodal fu- sion,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Dynamic multimodal fu- sion,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.440593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.946140Z digest=sha256:084fd5188ec15aac7ee16346c5c7ff3cb38b64fd6e682d3b0544b32a0063c659

Observation 725afab5-e854-41c6-a824-0e55ab6f22e1 · outbound

This paper cites Token-level contrastive learning with modality-aware prompting for multimodal intent recog- nition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Token-level contrastive learning with modality-aware prompting for multimodal intent recog- nition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.427623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.949590Z digest=sha256:c85017d2d3e9b684193ce84a2a96e3675077ae6d1c478afc233b27ed2104ee1d

Observation f6560aa6-3759-4685-af4e-8dd0cb368bbd · outbound

This paper cites Factorized Contrastive Learning: Going Beyond Multi-view Redundancy.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Factorized Contrastive Learning: Going Beyond Multi-view Redundancy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.953000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.953000Z digest=sha256:83339c70ecdea9f6fca7c22fb1c1b84b12867ef8e3b66ed34678d8b9c338fb43

Observation 1df649bb-1446-4120-9b21-e40bca68dbd0 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Representation Learning with Contrastive Predictive Coding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.957027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.957027Z digest=sha256:dc3b4f120be0cd9b9356305570f5a611938ea08477aeb4988da2bdbe218ca80d

Observation 6b1392ee-69b7-4f76-9a83-7fa95d545827 · outbound

This paper cites Mag-bert: Multimodal adapta- tion gate bert for multimodal sentiment analysis,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Mag-bert: Multimodal adapta- tion gate bert for multimodal sentiment analysis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.415072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.960653Z digest=sha256:d5238b4275c36da6f151a9d8dce17abcee91bcffd7f561d7e1e94db880f60cf3

Observation d6829b96-a73a-483e-b91e-3d4e67226f81 · outbound

This paper cites Super- vised contrastive learning,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Super- vised contrastive learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.402249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.964160Z digest=sha256:099fb223454d62f5f7f12ebcbe2964ea1e8116d2ce3d95008560345859554b3f

Observation 21c51923-216d-4caf-8526-ffb2664bc9f1 · outbound

This paper cites Contextual augmented global contrast for multimodal intent recog- nition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Contextual augmented global contrast for multimodal intent recog- nition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.389489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.968055Z digest=sha256:12c522bbc0c0f86fb15ff7f2bb08d648e99da3daad0c787a5ea699036161cc4b

Observation 4bc8d6e4-68b3-439c-acc2-7c90437a8740 · outbound

This paper cites Prototypical net- works for few-shot learning,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Prototypical net- works for few-shot learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.375484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.971689Z digest=sha256:0291961127a6f45b7c9f3045fdc6c01d4e2940845e934428f5e13ce3c6d0ffbc

Observation 784331fc-9e48-4d4a-997b-281675232a6b · outbound

This paper cites Mintrec: A new dataset for multimodal intent recognition,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Mintrec: A new dataset for multimodal intent recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.358934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.975806Z digest=sha256:b997e9fcb8e3c59e5d6f387b2396c69ced148597f07dd28cd4bd67949cc24f0d

Observation 933be9cd-5da7-4c99-a6c3-6a247770e382 · outbound

This paper cites Mintrec2.0: A large-scale benchmark dataset for multimodal intent recognition and out-of- scope detection in conversations,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Mintrec2.0: A large-scale benchmark dataset for multimodal intent recognition and out-of- scope detection in conversations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.345706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.980139Z digest=sha256:423783cd27df589700a4db91743c1d48906fe61ee12bd6bdb4b9104213972637

Observation 03986d9a-d258-4798-b590-7c29da0d16be · outbound

This paper cites Understanding con- trastive representation learning through alignment and uniformity on the hypersphere,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Understanding con- trastive representation learning through alignment and uniformity on the hypersphere,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.332076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.984268Z digest=sha256:bc5f66671776e922f1f0169a8cd2866c0279fc5eb286abca493c1723c288c2ab

Observation d07c04cb-57af-4065-ab7c-fe834c1ae367 · outbound

This paper cites Can contrastive learning avoid short- cut solutions?,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Can contrastive learning avoid short- cut solutions?,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.318269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.988416Z digest=sha256:7877209a9ceb00d44d665f9d330e9855bd00d9271fb3adee5c3bfb6bba886aa9

Observation 82b0a686-7890-495e-8ac5-d004893d7a58 · outbound

This paper cites HiCLIP: Contrastive Language-Image Pretraining with Hierarchy-aware Attention.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion HiCLIP: Contrastive Language-Image Pretraining with Hierarchy-aware Attention

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.992396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.992396Z digest=sha256:993215a64712608a1e27716147dae14dcdf80d7c38db171a41ee77e65d5958ec

Observation ec593fda-efe6-4013-920b-613fda3256ba · outbound

This paper cites Attention bottlenecks for multimodal fusion,.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion Attention bottlenecks for multimodal fusion,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:51:47.304860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T15:51:46.996181Z digest=sha256:3c76eb2fff538e02361545b13e93c9e61611db35b1c85d1b5f5119df604018e9

Pith citing papers

Observation f528d143-3c65-4591-bbea-a709f30617cc · inbound

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion cites this paper.

MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:51:46.891087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:51:46.891087Z digest=sha256:94d6bad2d0dc2ef3e9da269dad6e2169bcc466b692923a7c91bf2528170caa8c

Observation b3106e20-e87d-4058-8b4d-fc168df99611 · inbound

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding cites this paper.

Modality Agreement- and Conflict-Aware Prototype Hypergraph Learning for Multimodal Intent Understanding MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-08T00:49:09.580828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-08T00:49:09.503196Z digest=sha256:20bf05be57f2334bb0b32feaf24f3b0d4b2141d0781e18f74e90f375670baa4b