Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:08:58.980743Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2507.11892.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:08:58.980743Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2f80eff5-5e07-4e4f-bc11-824bde58144a · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition The extended cohn-kanade dataset (ck+): A complete dataset for action unit and emotion-specified expression,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8f33e461-9f0b-45e0-87f3-fc36ada3e9af · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Induced disgust, happiness and surprise: an addition to the mmi facial expression database,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c7d2cc0f-ba65-45e7-82ba-d384efa2e394 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Facial expression recognition from near-infrared videos,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 697b9dc7-4c81-4042-ba30-c4b59b5cdca8 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Dfew: A large-scale database for recognizing dynamic facial expressions in the wild,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7b370099-2a7c-449f-acc8-6c8f058cd5f8 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bbf98089-a486-4f67-9f52-4f89fe49eff0 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Ferv39k: A large-scale multi-scene dataset for facial expres- sion recognition in videos,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 154984c9-e876-47c2-84ae-4d67398609e4 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Dep-fer: Facial expression recognition in depressed patients based on voluntary facial expression mimicry,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 770416af-30bd-44e0-ab48-7554f99f46f0 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Efficient facial expression recognition with representation reinforcement network and transfer self-training for human–machine interaction,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation afccfec2-b81d-41ca-baa9-d4b7a9b331e4 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Predicting personal- ized image emotion perceptions in social networks,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e3d8a74-1cb2-45fd-9683-5afaff8dc945 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Spatio-temporal convolutional features with nested lstm for facial expression recognition,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5639bf19-85ab-4294-8870-a9829cdffe31 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Saanet: Siamese action-units attention network for improving dynamic facial expression recognition,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fe8fd94-4f81-4457-8460-8a451088566b · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Deep residual learning for image recognition,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76ef193a-40c2-48a4-8ada-3bf341c47643 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Former-dfer: Dynamic facial expression recog- nition transformer,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80ec1c43-07f3-4a26-b909-4f2dc26d0eb7 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Ex- pression snippet transformer for robust video-based facial expression recognition,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b439d16-a4e4-443e-8a0d-f2d8140ab652 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Freq-hd: An interpretable frequency-based high-dynamics affective clip selection method for in-the-wild facial expression recognition in videos,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18c7a973-d1b5-4155-985c-bc9ab4855e05 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Facial expression recognition with adaptive frame rate based on multiple testing correction,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a028bf9e-5a31-4acf-b5d9-cbfd4792e5d4 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Video-based facial micro-expression analysis: A survey of datasets, features and algorithms,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7fd07951-24e8-4737-90d0-a71bd023e60d · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Dynamic facial expression recognition under partial occlusion with optical flow reconstruction,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c2d87dcf-a356-49e0-8c39-70c1a6bfed9d · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition From static to dynamic: Adapting landmark-aware image models for facial expression recognition in videos,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9feacc7-04c8-4699-b675-a01cc7560e19 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition A Survey on Facial Expression Recognition of Static and Dynamic Emotions
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e0c819f-6ba6-426b-852d-26ab597a8c06 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Prompting Visual-Language Models for Dynamic Facial Expression Recognition
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e4d0b94-72eb-4c57-bbbd-9f175b5202df · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Emoclip: A vision-language method for zero-shot video facial expression recognition,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 964eea29-3d44-4c2b-b79f-40a7c4518bee · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Domain knowledge enhanced vision-language pretrained model for dynamic facial expression recognition,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfdaf705-c9c9-4083-b02b-5a5371ec2dfe · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Enhancing zero-shot facial expression recognition by llm knowledge transfer,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 65ef274e-e86c-4c71-a2ac-6d5b54a12094 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Finecliper: Multi-modal fine-grained clip for dynamic facial expression recognition with adapters,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1013a341-58d0-45a2-ba81-9eef0ddfdd18 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Describe your facial expressions by linking image encoders and large language models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4efb4a3a-9c7f-49e8-8ad4-233405174190 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Bert: Pre-training of deep bidirectional transformers for language understanding,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acbe33dc-ae8f-43e1-a5a4-b2d18cd43e08 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Hierarchical Transformers for Multi-Document Summarization
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 725a3fc7-35fd-49bc-ac14-a37a4d0e6b46 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition HDT: Hierarchical Document Transformer
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce50ff7e-9644-4481-b55c-2d587c8e3144 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Multi-task learning of hierarchical vision-language representation,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0ca66b65-6b30-4633-b844-1cc8f35be18f · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 242d3e97-996a-4f4f-86b9-67bcf9fb00d2 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Hierarchical modular network for video captioning,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb83ea55-fd32-41b7-87bb-c41d4f594515 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Learning transferable visual models from natural language supervision,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 190d69ec-35eb-46a9-8f1f-3895e2b534c9 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Ceprompt: Cross-modal emotion-aware prompting for facial expression recognition,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 06cef10d-69f3-47cb-a657-62d3cab51910 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Sinkhorn distances: Lightspeed computation of optimal transport,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5110dfd2-d14c-4c5c-9502-d00fe488b898 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Recent advances in optimal transport for machine learning,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7cb898f7-e789-4677-a77b-f6aee9e2dcfe · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Reliable weighted optimal transport for unsupervised domain adaptation,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff886e30-8ed9-46c3-8ff7-4223c9bcf03d · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Unsupervised learning of visual features by contrasting cluster as- signments,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8550292e-d670-4762-98d3-edd840d6d036 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Optimal partial transport based sentence selection for long-form document matching,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee3bd80c-b722-45b9-8bfb-73909537e983 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Learning to align sequential actions in the wild,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55dcb424-4071-4d6f-abbd-814214ae3886 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition What when and where? self-supervised spatio-temporal grounding in untrimmed multi-action videos from narrated instructions,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a0a922a5-baee-4e19-869d-6cc760f1c08f · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Multi-granularity Correspondence Learning from Long-term Noisy Videos
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3acb2635-48d1-4bd5-8d6f-82a43a92f8d2 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Spatial- temporal graphs plus transformers for geometry-guided facial expression recognition,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 22258bd3-e252-4538-b727-1377306c1864 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Multimodal transformer for unaligned multimodal language sequences,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3b68206-aaa7-440c-8809-87610c2582a6 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Vilt: Vision-and-language transformer without convolution or region supervision,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adc6628f-859d-410f-b0bc-9ed5bdce20f6 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91adfa9d-5d95-413a-b0dc-543745217bdc · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Cliper: A unified vision-language framework for in-the-wild facial expression recognition,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a771a906-1941-46cd-9384-d776e716758d · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ef3652f-b4ba-4a6a-9be8-6020b5bdcb3b · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ff0bbb6-eb8a-41d8-ba58-e1eb879b040d · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7fdddc8-f420-4089-a31d-40ca86630da1 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Videomae: Masked autoen- coders are data-efficient learners for self-supervised video pre-training,
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0f13447-3841-4612-bf12-eab2e9546562 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Mae-dfer: Efficient masked autoencoder for self-supervised dynamic facial expression recognition,
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40c53bc4-6176-455b-b879-1db2e02025b8 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image cap- 14 tioning,
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d0349cd-faf4-4122-9945-7563471fbdbb · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Frozen in time: A joint video and image encoder for end-to-end retrieval,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3208862-0942-423d-9cbc-8b85e139500c · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Emotion recognition using imperfect speech recognition,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0edd08c9-2483-4ae2-b727-2e19eecae091 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Posterior calibration for multi- class paralinguistic classification,
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b06a9716-7b97-40fa-a894-d07f957a2a4c · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Rethinking the learning paradigm for dynamic facial expression recog- nition,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9fc0632-63cf-4362-9066-95f90e68fbd1 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition A$^{3}$lign-DFER: Pioneering Comprehensive Dynamic Affective Alignment for Dynamic Facial Expression Recognition with CLIP
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb674402-af34-407a-86c6-4a7b01908490 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Clip-aware expressive feature learning for video-based facial expression recognition,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c53bea4-72b1-4926-8ec7-ed45a3ceb9b8 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition NR-DFERNet: Noise-Robust Network for Dynamic Facial Expression Recognition
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c01291c4-fcdc-4459-8ccd-dd3fbabce61f · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Logo-former: Local-global spatio-temporal transformer for dynamic facial expression recognition,
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 025acfcc-723b-4024-9ea7-4130c4153e70 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Intensity-aware loss for dynamic facial expression recognition in the wild,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 44e6292c-dc8c-490b-8ef8-c2126613213e · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Transformer-based multimodal emotional perception for dynamic facial expression recogni- tion in the wild,
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d6c81c3-f280-4671-9ef9-d9f219746c94 · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Svfap: Self-supervised video facial affect perceiver,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56f4fb15-b13d-4326-8e4b-6bd0070b72bd · outbound
From Coarse to Nuanced: Cross-Modal Alignment of Fine-Grained Linguistic Cues and Visual Salient Regions for Dynamic Emotion Recognition Hicmae: Hierarchical contrastive masked autoencoder for self-supervised audio-visual emotion recogni- tion,
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.