Pith. sign in

Paper Citation Record · LEDGER

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions

As of 7 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 0 inbound Pith citation observations for arXiv:2507.21015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21015 v1

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:07:27.714660Z

measured 94 of 94 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

94 of 94 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ef8048ad-0f81-48e2-95bd-6228b7193da3 · outbound

This paper cites Society of mind.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Society of mind

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.592232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.592232Z digest=sha256:3326456bd2bf56472f304d3f00041912a75ac6a44b75391863e4525026851367

Observation 83c82526-bf1a-4ee5-a186-383f493b7c67 · outbound

This paper cites Emotion recognition in human-computer interaction.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Emotion recognition in human-computer interaction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.661748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.661748Z digest=sha256:5a9c4fb2d3e5c971504f8c4b5dbd070de34f26628bc6ddf17a834b3572e0b4e6

Observation f491c3aa-26ad-479e-9029-d58d3660f428 · outbound

This paper cites An overview of emotion in artificial intelligence.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions An overview of emotion in artificial intelligence

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.741061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.741061Z digest=sha256:5233f4daf0515f2bf30520a9dbbfad609235b70a1822afb25f15147d0ca1914c

Observation e807f6d5-3610-46d3-b508-83399eee5a45 · outbound

This paper cites Deep facial expression recognition: A survey.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Deep facial expression recognition: A survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.793886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.793886Z digest=sha256:5b75b535afe4fa5aa32d5575220d967180970472b2da2cf0db65405ff39e13e8

Observation 1dba552a-29c9-4621-8fd8-d36930e6142a · outbound

This paper cites A survey on facial emotion recognition techniques: A state-of-the-art literature review.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A survey on facial emotion recognition techniques: A state-of-the-art literature review

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.873360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.873360Z digest=sha256:8808868eda75a226c2c86730e384fe3186a48f4f61707aa4d49edf155e0924be

Observation a0818c7a-b929-4123-a7ab-abe8debb2fcc · outbound

This paper cites Understanding deep learning techniques for recognition of human emotions using facial expressions: A comprehensive survey.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Understanding deep learning techniques for recognition of human emotions using facial expressions: A comprehensive survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.940239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.940239Z digest=sha256:b96ec85bfb754d9421f1e20a9ad8b7d44166988918f9d7016faecc0aeccbcbce

Observation 1a786f0b-de19-4653-b951-d668d2a57ee0 · outbound

This paper cites Facial micro-expressions: An overview.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial micro-expressions: An overview

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.056038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.056038Z digest=sha256:71792c2a175bfc2ae9f9147122734d9021ea41fad9e231fd0253c68eaaf5e35a

Observation 2c240e3f-d6e6-4c7d-abab-3ca7006dbceb · outbound

This paper cites A model of the perception of facial expressions of emotion by humans: Research overview and perspectives.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A model of the perception of facial expressions of emotion by humans: Research overview and perspectives

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.176560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.176560Z digest=sha256:ef6f1a97b87eef6c2ed8e86089cf2adbdc49537ef92ebce28c962f0db04271d3

Observation d97ce718-b408-409d-902e-c023dde91c0f · outbound

This paper cites Deep learning for human affect recognition: Insights and new developments.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Deep learning for human affect recognition: Insights and new developments

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.230927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.230927Z digest=sha256:f326f902f00f54865682ca5c8e98b49252b63306ff292f2f44d00f5360b9fa23

Observation dd551dc5-478a-40a1-ac56-ad3a6f194767 · outbound

This paper cites A review of affective computing: From unimodal analysis to multimodal fusion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A review of affective computing: From unimodal analysis to multimodal fusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.295239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.295239Z digest=sha256:ef2ba37bb78bf3a80dfbff6c7ef28183a13e9bd2c87e529c19bf62bd1b03a72f

Observation a13612c8-151f-425a-bf05-fd6c844a98a8 · outbound

This paper cites An argument for basic emotions.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions An argument for basic emotions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.299794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.299794Z digest=sha256:e2414881ab4cfe15c1a8739ec8731f21729ed32de8654f9d479caa52ce3277be

Observation e43912a0-793e-476a-ab88-55e0bc60b856 · outbound

This paper cites A circumplex model of affect.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A circumplex model of affect

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.304830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.304830Z digest=sha256:8d48e2922f99017d02c52427207ab723b8656a4db1bd42c8ed02f499d8457647

Observation b75452e2-5d5f-492f-b21f-126db4a70919 · outbound

This paper cites OV-MER: Towards Open-Vocabulary Multimodal Emotion Recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions OV-MER: Towards Open-Vocabulary Multimodal Emotion Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.309698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.309698Z digest=sha256:d174786faff4146f90a42873172d92434e0cd64eaf66106e06b3b13eada83565

Observation 8450f0b3-3ee7-4861-a546-d95f9703d13b · outbound

This paper cites GPT-4o System Card.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions GPT-4o System Card

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.314585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.314585Z digest=sha256:7e58fc008f50192ba1dd6f6a741b0fa280b7582ceba9975034172868a16b211c

Observation 3f4eb40c-0771-4fde-b84e-45e9ef730be2 · outbound

This paper cites Self-report captures 27 distinct categories of emotion bridged by continuous gradients.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Self-report captures 27 distinct categories of emotion bridged by continuous gradients

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.320211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.320211Z digest=sha256:21171f78217685a75f0f19b10264a02df8e837c47a1b2e574968d4300baeba47

Observation 8e58cf6f-20f1-4b0b-bed8-4ba023332c5e · outbound

This paper cites The language of emotion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions The language of emotion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.324876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.324876Z digest=sha256:990084016cac387e114e77763fb4a30dcfb36277ced3c0eb238433631fe55ab8

Observation 474ccc17-737c-4a19-b8e1-b08742ecb011 · outbound

This paper cites The role of language in emotion: Predictions from psychological constructionism.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions The role of language in emotion: Predictions from psychological constructionism

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.190794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.329266Z digest=sha256:16af843cc17638c280861a936be526f3752cc9136e359164a41b2be0d92e4d74

Observation 613824fc-1400-4cf0-ae0d-3f896e0f5bce · outbound

This paper cites Describe your facial expressions by linking image encoders and large language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Describe your facial expressions by linking image encoders and large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.174981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.333575Z digest=sha256:1074001f711a3e983be3858ef96aa247ede6d984a15955b39fd7df184071a2e6

Observation 6747419b-2eb9-4c31-8d2e-fbce8c61250f · outbound

This paper cites Facial affective behavior analysis with instruction tuning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial affective behavior analysis with instruction tuning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.158768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.337588Z digest=sha256:5d5694ff9669097805f7b07774e087c7c4ff651a8c09f6baf64e5524400eea8f

Observation 00e97ee9-0119-45ec-8c4d-9cfcaf6414f9 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning transferable visual models from natural language supervision

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.342458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.342458Z digest=sha256:4bc1ad0af3c6813007907ce80bcc1504a7d26be43619367f78b896af80ad4e75

Observation 27759557-c862-47df-b6a3-0975c4adcc68 · outbound

This paper cites Emoclip: A vision-language method for zero-shot video facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Emoclip: A vision-language method for zero-shot video facial expression recognition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.132614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.346762Z digest=sha256:c43722f16c7b99dd44b8093dff603eaec136a87ebf6156f7b96328043ec98d4e

Observation 63b2af78-2d5c-48e6-a36e-d6e31ff541ff · outbound

This paper cites Flip-80m: 80 million visual-linguistic pairs for facial language-image pre-training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Flip-80m: 80 million visual-linguistic pairs for facial language-image pre-training

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.117499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.350761Z digest=sha256:46c4a10bfe1b5d09192df25979391f349c10101e8dabb9317af548c5c957ee68

Observation af6ee1a3-6628-434e-8d7d-7bacdc929a0f · outbound

This paper cites Enhancing zero-shot facial expression recognition by llm knowledge transfer.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Enhancing zero-shot facial expression recognition by llm knowledge transfer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.102546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.355611Z digest=sha256:390013d7fe6752f5355675f2d42a6dba52540f55cdb142fe7e0c4c2ef949f603

Observation 796b78a5-1e88-4291-95cf-94103d0cd1ba · outbound

This paper cites Facexbench: Evaluating multimodal llms on face understanding.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facexbench: Evaluating multimodal llms on face understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.360923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.360923Z digest=sha256:373368797597cc2c5f5475e7d6f7165d80622780e627bce109990ea5ddeee315

Observation 101f4a02-2524-4b60-84a9-2fee07ea7bc7 · outbound

This paper cites Face-human-bench: A comprehensive benchmark of face and human understanding for multi-modal assistants.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Face-human-bench: A comprehensive benchmark of face and human understanding for multi-modal assistants

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.365402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.365402Z digest=sha256:42863728a9f31c734a6739174a9ff2d178bb88ea049cda2282507a44bbf76019

Observation 9093d8a4-4c21-4974-bc0d-27f203cb14d5 · outbound

This paper cites Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.370480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.370480Z digest=sha256:bc895797cc8e876fa4ea519bee764206b675d1470886bb12a3093f5883fe59eb

Observation 3b546bfb-8e31-4df9-9027-375b17c4b38a · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.375535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.375535Z digest=sha256:4b48923947abf908304642be2ad630c13608f50fb10c6808d9e1a37ddd1ce522

Observation 0d4882c9-0f73-42ab-ba71-9383e160156d · outbound

This paper cites Occlusion aware facial expression recognition using cnn with attention mechanism.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Occlusion aware facial expression recognition using cnn with attention mechanism

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.075874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.385225Z digest=sha256:f583130fd0ae2bf1a7f59d137881e7b082af9a6eb3052814b8c001ab3276edf8

Observation c26d7ed2-1a9b-4b0f-beab-c05733330d07 · outbound

This paper cites Region attention networks for pose and occlusion robust facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Region attention networks for pose and occlusion robust facial expression recognition

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.061244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.390540Z digest=sha256:3855505946728e778661131aee4bc8559984d900a817b3c38505ef1470f04bc4

Observation 430b144b-c4bd-4e05-800f-4dd2e61dbb53 · outbound

This paper cites Learning deep global multi-scale and local attention features for facial expression recognition in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning deep global multi-scale and local attention features for facial expression recognition in the wild

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.044843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.395215Z digest=sha256:502847b3a2928739744905fc9c98fe11e5450dc45a0f0aa517164df9fdfda386

Observation 7e04cd3e-a95e-4130-a8b5-01fc21efedf3 · outbound

This paper cites Robust lightweight facial expression recognition network with label distribution training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Robust lightweight facial expression recognition network with label distribution training

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.029929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.399715Z digest=sha256:3900fccc90ec51b3427c014a7e3357adcf0274a3353a62b4d8859a04d31b4440

Observation 80dd1b37-590b-4991-b264-f22b09697a1a · outbound

This paper cites Facial expression recognition with visual transformers and attentional selective fusion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial expression recognition with visual transformers and attentional selective fusion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.014568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.407722Z digest=sha256:5604545680e06a6ea1ff19ea52599d3f6cf4e163646543086875899ec8cc2c4b

Observation c1ea6df3-43d5-425c-a3df-94839e7e39de · outbound

This paper cites Transfer: Learning relation-aware facial expression representations with transformers.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Transfer: Learning relation-aware facial expression representations with transformers

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.998550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.411851Z digest=sha256:f0b8ee5573ea93bc7854a51c68aea4daa38d4a108e757ff7548bf66b7a68929a

Observation 4dbb4180-0854-43b3-a52d-afac2bfb917e · outbound

This paper cites Poster: A pyramid cross-fusion transformer network for facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Poster: A pyramid cross-fusion transformer network for facial expression recognition

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.979206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.415946Z digest=sha256:e5e329aee09ef7f1c8ae1ce726fbd409452c33ad43ef6fbbe9026e383971b854

Observation 334c6763-05d5-4a4d-b5c1-ded126b9604a · outbound

This paper cites Svfap: Self-supervised video facial affect perceiver.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Svfap: Self-supervised video facial affect perceiver

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.963118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.420289Z digest=sha256:3edcffc1d8a19e3f283a434678cda954f0da0c884fe016b1d671486fc8a398a0

Observation b831ece9-cd9f-4e18-9f18-1e7875b8a1ed · outbound

This paper cites Poster++: A simpler and stronger facial expression recognition network.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Poster++: A simpler and stronger facial expression recognition network

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.943246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.424789Z digest=sha256:70c8f753933c9c2506ff673077c65655798e53b052a921e3d6ca8a25794d1d4b

Observation 5d933e14-0c5a-4771-8e24-f997f8411e73 · outbound

This paper cites Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.927537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.431081Z digest=sha256:174034cbf3574eecfad8fb628f38b7f6768fcf1a93015959be21ef90218507b3

Observation d4b4681a-434e-4c5f-89b7-9ff5626202a1 · outbound

This paper cites Suppressing uncertainties for large-scale facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Suppressing uncertainties for large-scale facial expression recognition

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.910444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.435495Z digest=sha256:75c5fe9c934cc259eb82f4c1801baab2631df0168f738eff2ea7a4da7e0f29bd

Observation d9df4532-1db3-4df1-9c92-b1a861982c28 · outbound

This paper cites Relative uncertainty learning for facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Relative uncertainty learning for facial expression recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.892265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.439812Z digest=sha256:9c797a2f358316cff4d3469028cc3560b5327d7b3723ff10080cfda51c22208e

Observation ec11e3d6-aea9-44c9-9e48-c902aa3a5b2a · outbound

This paper cites Learning emotion representations from verbal and nonverbal communication.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning emotion representations from verbal and nonverbal communication

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.874331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.444071Z digest=sha256:b8d21e518382650faeca54e0889f9728c6c0b6a8c026566bb894714a25dff310

Observation 668b0632-e627-4b26-ac6b-6ece05f3dba8 · outbound

This paper cites Collecting large, richly annotated facial-expression databases from movies.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Collecting large, richly annotated facial-expression databases from movies

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.857788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.448368Z digest=sha256:d90618fa2761631a14d43f45c5cc167403f30d3fc014671ed7703e3cbe2bfd63

Observation 53205689-aba6-4690-b331-ea5483dfff10 · outbound

This paper cites Training deep networks for facial expression recognition with crowd-sourced label distribution.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Training deep networks for facial expression recognition with crowd-sourced label distribution

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.842244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.453976Z digest=sha256:c4c2c5114bcb4335d3211ec751de513191f34806fdbf4f9af5329bcc5e96b5ca

Observation d9730ec3-50fe-4659-b3d4-074cc7ab87f4 · outbound

This paper cites Affectnet: A database for facial expression, valence, and arousal computing in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Affectnet: A database for facial expression, valence, and arousal computing in the wild

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.825975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.458964Z digest=sha256:3342dffd916a2c666a6a696da79c05b47dcce81c0a2a79672078c66dfb77d7d0

Observation 2b92e18b-5709-42bd-8711-922081a8989c · outbound

This paper cites Dfew: A large-scale database for recognizing dynamic facial expressions in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Dfew: A large-scale database for recognizing dynamic facial expressions in the wild

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.810550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.463896Z digest=sha256:aa16a9598b69d55a0e3ec56e5f311c84425c158192837b89efbe2736568ef979

Observation e5d8ca84-2fa9-4848-bfc6-bfea2b3c7214 · outbound

This paper cites Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.794142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.468537Z digest=sha256:e5ff53f78b8db1c4bee3eccbc933a96fecc6524cfb5fd50e78d095df8c9d5089

Observation 859685ba-3115-451b-8b6d-2e6b1be7dad7 · outbound

This paper cites Ferv39k: A large-scale multi-scene dataset for facial expression recognition in videos.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Ferv39k: A large-scale multi-scene dataset for facial expression recognition in videos

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.777832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.472829Z digest=sha256:74f335925aca9dcdae9359db23b028f41bfed3bfc4e50565ba7d0666134f9de0

Observation a6b0f5c3-de36-4929-8dee-e00c386c8737 · outbound

This paper cites Aff-Wild2: Extending the Aff-Wild Database for Affect Recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Aff-Wild2: Extending the Aff-Wild Database for Affect Recognition

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.477365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.477365Z digest=sha256:623cd228ce9b3c887a2cc24db73aa65948bde62690342f1b54ff5ae1d59b138f

Observation 45a3adf5-cf83-46dc-adcf-6d8fe902d8c9 · outbound

This paper cites Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.760052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.482389Z digest=sha256:738d5dc078411f72d05b4c386fd173a09787df087f4e053b25c03ddbf378a9b9

Observation 68a872fe-c5e7-4beb-b4dc-a0bd6f360814 · outbound

This paper cites Compound facial expressions of emotion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Compound facial expressions of emotion

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.742804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.486532Z digest=sha256:3d1db948b24498e06110120b1072216f9fdcca2623c585dbebea52d46b61d3ac

Observation bfcf1974-de71-402c-b1eb-b7261af5129e · outbound

This paper cites Configural information in facial expression perception.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Configural information in facial expression perception

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.727222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.491038Z digest=sha256:117927bd0481cfd1e0790b7b3dfc6b99a20cae86306d275403c49a63eedf4f62

Observation 9c4b2c25-bfde-4039-a2a1-ca4eb145d015 · outbound

This paper cites Parts and wholes in expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Parts and wholes in expression recognition

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.712447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.496626Z digest=sha256:b96ed1caf3f0a45bfc03374fc19fbcb61c7f9593921e51423a089879cfc5f43f

Observation 736555e8-e085-4135-9928-40a153ecaa55 · outbound

This paper cites Mixed emotions: Holistic and analytic perception of facial expressions.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Mixed emotions: Holistic and analytic perception of facial expressions

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.696994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.501743Z digest=sha256:78c4d22d301556d5c6edb5114e4ac934db0237d1e186e0e59610117f915226df

Observation c168dbc8-1d94-4a55-baaf-e90bf4d73d38 · outbound

This paper cites The role of facial movements in emotion recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions The role of facial movements in emotion recognition

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.680641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.506388Z digest=sha256:cdb4ccc639a54c45b6939d9c47e1142aa312f489cbcb859a6c6803b2d3b6488d

Observation cc17857f-4354-4c3a-afae-b0d9837b3146 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.510794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.510794Z digest=sha256:0b001140138f93e9eb49c52ecf6569f48ce55dead29bb62c780e535a7fa2f651

Observation 45b1d21e-5166-4104-b63c-945dd671881e · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Scaling up visual and vision-language representation learning with noisy text supervision

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.515559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.515559Z digest=sha256:bfff5b16205f209219074d65383b919e8fa93014d3df1735d65d84d67bf1d761

Observation 543b6e97-02fa-484a-8e72-b90f188ba575 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Reproducible scaling laws for contrastive language-image learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.519813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.519813Z digest=sha256:0753d71bfc3b3e6f00a0184ce0715507ed127c937f11829ca4e5b4cd6babdef1

Observation 16766737-fb2d-4040-aa01-201cb94a8304 · outbound

This paper cites Sigmoid loss for language image pre-training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Sigmoid loss for language image pre-training

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.524336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.524336Z digest=sha256:9ff329b979ad39b41bca03f6d16a7ab39248fc26d561eadf98b59fa831b49d6f

Observation a58c779b-930f-4c26-a463-bf3d2c91a56a · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.528659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.528659Z digest=sha256:c1492ec0f0becc124dbf1adc5a219bf1ee08076cbc49f243310cf81b48850235

Observation e367f1a0-e1a6-41ea-b20d-c15f86c97ebf · outbound

This paper cites Demystifying clip data.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Demystifying clip data

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.621180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.534253Z digest=sha256:ae7835b6b7260aaca6be1b551c134fd932b49216e3bf5aebc4837321d558e205

Observation 296cdcf5-98d0-4dd3-a98d-6a1461eba2ee · outbound

This paper cites Dreamlip: Language-image pre-training with long captions.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Dreamlip: Language-image pre-training with long captions

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.605700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.539781Z digest=sha256:fae5d5404dfedf9d864f01688a8dfa8f5cb670ce9eebc34d424fc30d45da2ad2

Observation 172bed96-dc4c-4ace-b307-44b8a6db6d9e · outbound

This paper cites Modeling Caption Diversity in Contrastive Vision-Language Pretraining.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Modeling Caption Diversity in Contrastive Vision-Language Pretraining

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.544588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.544588Z digest=sha256:e22d5e0a96caad1e6dacad26823950f52f05443a2bd0f9125c8e64710fe130b9

Observation 96e16bd9-f574-4353-8097-d51cf7a79905 · outbound

This paper cites Improving fine-grained understand- ing in image-text pre-training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Improving fine-grained understand- ing in image-text pre-training

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.590425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.549311Z digest=sha256:77184f2336f2efe736657361bb5e0bcd0faa86452ed96c823cbebe09c4836dfd

Observation 2a162f3a-c265-4461-bf56-c969dce59c59 · outbound

This paper cites General facial representation learning in a visual-linguistic manner.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions General facial representation learning in a visual-linguistic manner

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.554743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.554743Z digest=sha256:29844882920d6f1e4c3bfc048f0889001a401bec113b06b98951074cf157fb51

Observation d1da1679-01f1-4495-bf09-5f9664c2e4ca · outbound

This paper cites GPT-4 Technical Report.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions GPT-4 Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.559162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.559162Z digest=sha256:7d2c9195c58052fd6a7fe8fa1c864dae5fe0841d8e0489f6bf41094e2bc9dcf5

Observation 62611581-b192-4cca-832b-9a3429ff96ba · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.564018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.564018Z digest=sha256:df425d2931cd19795e8b12498c030dc687c56b8de8f7d3cb0942b4a2507c5f5f

Observation 18510aaf-16b3-425d-8e0a-92146c5d651d · outbound

This paper cites Minigpt-4: Enhancing vision-language understanding with advanced large language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Minigpt-4: Enhancing vision-language understanding with advanced large language models

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.553765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.568575Z digest=sha256:fc94410a36077fb0001097c215bea9f1ea569d17569bf46057111224c42d6413

Observation a7fe66eb-2834-4485-abe6-a6200aecf3bf · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.573362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.573362Z digest=sha256:0cf354ff4d03760da539d9c35a900269e762cc2b89319ec041a4e528878af641

Observation 0846c629-f9e1-45ee-94ff-fbfa015973c2 · outbound

This paper cites Qwen Technical Report.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Qwen Technical Report

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.577897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.577897Z digest=sha256:1771d4b08c663c97e5d408ad90dfddefd1b4ce8fb6eb35d30cfab5962a05a23b

Observation b32689e9-eb62-4a25-9e74-442c05fcc889 · outbound

This paper cites Expllm: Towards chain of thought for facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Expllm: Towards chain of thought for facial expression recognition

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.527127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.583046Z digest=sha256:3da903f7c422d10920768bd0af9ab9198063c3ff64ebe6bb5493bb05a30e7850

Observation 91706ae5-009f-4165-9389-758e3f05e0f2 · outbound

This paper cites EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.587777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.587777Z digest=sha256:8fa1841aa740dbd509158fe71d360b3562c46a3d244b4119b0411cce945ddcf8

Observation ad3f9e85-5b32-4f95-bb8f-4e6f97606552 · outbound

This paper cites Emotion-llama: Multimodal emotion recognition and reasoning with instruction tuning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Emotion-llama: Multimodal emotion recognition and reasoning with instruction tuning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.510694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.592827Z digest=sha256:18c9f5f01792cf8fb13332e49d55ab7d4c10763b0633f34e544c90e7813f7c4d

Observation 9ff15cd6-87af-4345-9f2f-47c7635f8d50 · outbound

This paper cites AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.597574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.597574Z digest=sha256:11faa3e6de1852116e8d4358688cb41731c653f95186dddd1702afce1bbe677a

Observation dcd2a90e-57dd-493c-898a-6e66e0d1b598 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.602296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.602296Z digest=sha256:677b1448f83a51d4bac33b692536956350585faaf9982c8ece36e7c00f1ee27d

Observation 372f1b4f-0d2c-4d48-9411-8e24d1f8fb4e · outbound

This paper cites Generative adversarial network for text-to-face synthesis and manipulation with pretrained bert model.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Generative adversarial network for text-to-face synthesis and manipulation with pretrained bert model

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.493528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.607503Z digest=sha256:1068a646df4e07d6b366c0db718ca0baea68e8b5a080b528256158eaddcd28e4

Observation d620aab3-8df8-4240-bf39-ae514f932d14 · outbound

This paper cites Tedigan: Text-guided diverse face image generation and manipulation.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Tedigan: Text-guided diverse face image generation and manipulation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.611815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.611815Z digest=sha256:65aeb0011bdebd861c71169e822d6c706cb307b47f42caa39c6dda714cdeca83

Observation bd60d693-e329-40fb-874e-ffc93d0969ba · outbound

This paper cites Talk-to-edit: Fine-grained facial editing via dialog.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Talk-to-edit: Fine-grained facial editing via dialog

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.463190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.616335Z digest=sha256:329c7d60f01a370fc445b25b646fb326e793e39e597e56d1d1fc955faac8461b

Observation c1535cd7-1311-4ab3-a765-abb76371da81 · outbound

This paper cites 15M Multimodal Facial Image-Text Dataset.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions 15M Multimodal Facial Image-Text Dataset

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.621712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.621712Z digest=sha256:cffb16a61ffa8890d498fd5aec4a0d940ed121ec78fa73c4b6f8a4cadcc7667d

Observation 100cbd00-c651-4570-a040-0f104650aa21 · outbound

This paper cites Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.626612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.626612Z digest=sha256:41ca3b2d2eda7fe527ede83e212c6e5659d39c42f836752afbd350269e4afb3d

Observation b5516e37-14cc-4f24-95d8-7281e94524eb · outbound

This paper cites Recognizing emotion from facial expressions: psychological and neurological mechanisms.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Recognizing emotion from facial expressions: psychological and neurological mechanisms

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.447457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.632099Z digest=sha256:d2b08f24083db63317b8fde54e661420e954a644b42a60eae63b7eebe7c05605

Observation c4942f20-be8d-43ad-9381-17fd57fdd3f2 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.636802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.636802Z digest=sha256:27e1f3a58185332068f17c194ce04ae15e30c40aa3f22567f06eb953e5a71575

Observation 49e8119b-58e9-439e-9a72-0efcf4c17de3 · outbound

This paper cites Attention is all you need.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Attention is all you need

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.641491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.641491Z digest=sha256:de97c2cd3dc063d241eebba65282ca6bc1d67dd271c4dcd4b8f500aecaa9a1fe

Observation 316b960e-eb7b-47f3-b8bd-6f6fc7c5b4d3 · outbound

This paper cites Facial action coding system.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial action coding system

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.420990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.646353Z digest=sha256:1ae6e370f9f81f09fcb1b9f69265640898fe37ae973e93f9cd1f28248e49e8e3

Observation 00c31e87-2361-43ec-9cb8-440841c6c4aa · outbound

This paper cites Softclip: Softer cross-modal alignment makes clip stronger.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Softclip: Softer cross-modal alignment makes clip stronger

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.404940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.651724Z digest=sha256:2d3298e49a671075bd67c4cb1d16735602eefd343e82f0bd39e06e60d012c743

Observation c92ff191-88ce-47ea-a195-157a86b31f93 · outbound

This paper cites Cwcl: Cross-modal transfer with continuously weighted contrastive loss.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Cwcl: Cross-modal transfer with continuously weighted contrastive loss

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.388027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.658642Z digest=sha256:8503da28546fdd1aaed4f37d5a1dc3b47c03594b5d0af7ff127975bf801b6ae9

Observation c7333be1-82c7-4c11-ab1e-5c7e57932a2d · outbound

This paper cites Generalizable facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Generalizable facial expression recognition

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.371188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.664046Z digest=sha256:bb48ebe10911c550352583f3b20ec77303ae129960996c6a61fe0a16fc25b202

Observation c6b01c4a-bc33-4570-bd7d-25213998ac48 · outbound

This paper cites Flava: A foundational language and vision alignment model.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Flava: A foundational language and vision alignment model

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.351256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.669757Z digest=sha256:b089de8158ef6d2a48e040da72878bf790734c041e7a86fa5cbc20d44b7c9555

Observation 37e65246-652e-4bcb-92b5-e08e324a1669 · outbound

This paper cites Face-MLLM: A Large Face Perception Model.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Face-MLLM: A Large Face Perception Model

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.675297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.675297Z digest=sha256:8ca27219face9f51a4ced04a38a95345785a8adbd76aa2d324dc2bae0f0602e5

Observation 38ee7e1c-45b8-44fe-a6f2-2dbb1596612e · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.680305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.680305Z digest=sha256:2bb663f245d07166cc854d3ea45345e9b3127e4aaa0be6e51a55bc884c89b6e2

Observation 4d20077b-717a-4e68-84af-bfacd77fd4f5 · outbound

This paper cites Learn from all: Erasing attention consistency for noisy label facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learn from all: Erasing attention consistency for noisy label facial expression recognition

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.333222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.686607Z digest=sha256:604ca4f313652fb7344319804f059e182fa17a03c7c8051dab7f485759d048aa

Observation 0c7ce87b-05cb-44d0-a49c-09dfc5bd7168 · outbound

This paper cites Latent-ofer: Detect, mask, and reconstruct with latent vectors for occluded facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Latent-ofer: Detect, mask, and reconstruct with latent vectors for occluded facial expression recognition

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.314024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.692832Z digest=sha256:221cf503dc5db1e70e2eeca9afcce1fdf68cfe1140d321587a56108e9ab0de8d

Observation 0ddd7f09-5bdb-4044-92eb-3384ca30156a · outbound

This paper cites From static to dynamic: Adapting landmark-aware image models for facial expression recognition in videos.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions From static to dynamic: Adapting landmark-aware image models for facial expression recognition in videos

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.295881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.698844Z digest=sha256:a3bbf1f82dffce9691fa6b6535ed16f6c3d911fd1ea1858d7b55196f4ff64cc5

Observation 84411ab7-998a-4dea-b93e-156964029dea · outbound

This paper cites Videoclip: Contrastive pre-training for zero-shot video-text understanding.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Videoclip: Contrastive pre-training for zero-shot video-text understanding

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.277285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.705145Z digest=sha256:e5784ab2501f474a1361bbab9e6adfee6709b98718bf7b68721ab2ceeb74d0a3

Observation bb9fa76a-8359-49e1-a6b9-59488c37731a · outbound

This paper cites Expanding language-image pretrained models for general video recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Expanding language-image pretrained models for general video recognition

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.261598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.709988Z digest=sha256:9e7d8c83121c6d1e79d36a53e132b461d29b1fbbd2693eca1d829905f35d06e6

Observation 30c11dac-bc73-4cf0-8cae-94b246a1bbc4 · outbound

This paper cites Learning to prompt for vision-language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning to prompt for vision-language models

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.243764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T13:07:27.714660Z digest=sha256:f3dd786e4a59432537b28ca531f7e8061d73003f0e24ce40d47ac6914e57a529

Pith citing papers

No inbound Pith citation observations are available.