Pith. sign in

Paper Citation Record · LEDGER

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions

As of 18 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 0 inbound Pith citation observations for arXiv:2507.21015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21015 v1

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:07:27.714660Z

measured 94 of 94 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

94 of 94 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ef8048ad-0f81-48e2-95bd-6228b7193da3 · outbound

This paper cites Society of mind.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Society of mind

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.592232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.592232Z digest=sha256:18647cda6496225cc5540959d713983cc3a3a58a5d3391f2de6e95173b5fe0c2

Observation 83c82526-bf1a-4ee5-a186-383f493b7c67 · outbound

This paper cites Emotion recognition in human-computer interaction.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Emotion recognition in human-computer interaction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.661748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.661748Z digest=sha256:efcdac7980a2fad2b984a3249b4f70e3965b38366b3152115b8b4003df46ed40

Observation f491c3aa-26ad-479e-9029-d58d3660f428 · outbound

This paper cites An overview of emotion in artificial intelligence.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions An overview of emotion in artificial intelligence

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.741061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.741061Z digest=sha256:5a350dc656dfd8b943c3be4c63b7791f65fd0da95dba88b4358553e5b830abfd

Observation e807f6d5-3610-46d3-b508-83399eee5a45 · outbound

This paper cites Deep facial expression recognition: A survey.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Deep facial expression recognition: A survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.793886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.793886Z digest=sha256:1d5dcd103dbf59c25ab33f92e2088f619fed3d9f3405318e1b65d2eded8d8c2d

Observation 1dba552a-29c9-4621-8fd8-d36930e6142a · outbound

This paper cites A survey on facial emotion recognition techniques: A state-of-the-art literature review.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A survey on facial emotion recognition techniques: A state-of-the-art literature review

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.873360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.873360Z digest=sha256:61fa062fc7e9b0f1153f860c36eee6196952f4d29033738093c65b0eadfdc222

Observation a0818c7a-b929-4123-a7ab-abe8debb2fcc · outbound

This paper cites Understanding deep learning techniques for recognition of human emotions using facial expressions: A comprehensive survey.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Understanding deep learning techniques for recognition of human emotions using facial expressions: A comprehensive survey

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:26.940239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:26.940239Z digest=sha256:c73423eb15389f85c2c09acc891d9f0f4898897766f38de7f1d5cdcb410e3863

Observation 1a786f0b-de19-4653-b951-d668d2a57ee0 · outbound

This paper cites Facial micro-expressions: An overview.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial micro-expressions: An overview

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.056038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.056038Z digest=sha256:68af806bd393af95051a2796efcde1ddf4a3ff3729d33b0d35d1b22678fe0178

Observation 2c240e3f-d6e6-4c7d-abab-3ca7006dbceb · outbound

This paper cites A model of the perception of facial expressions of emotion by humans: Research overview and perspectives.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A model of the perception of facial expressions of emotion by humans: Research overview and perspectives

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.176560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.176560Z digest=sha256:7bc4d926f874c6f79415831937985481b25678e79deadac63cddee9f5f9480a4

Observation d97ce718-b408-409d-902e-c023dde91c0f · outbound

This paper cites Deep learning for human affect recognition: Insights and new developments.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Deep learning for human affect recognition: Insights and new developments

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.230927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.230927Z digest=sha256:0aa7683c104bc409575315163ba85a03901a42031da93aa3c316dbdf7703c008

Observation dd551dc5-478a-40a1-ac56-ad3a6f194767 · outbound

This paper cites A review of affective computing: From unimodal analysis to multimodal fusion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A review of affective computing: From unimodal analysis to multimodal fusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.295239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.295239Z digest=sha256:5f9ec341d559d566d47bbb5a5e3dde79c6d28e1d3f45f4105241862a5b1d9ce9

Observation a13612c8-151f-425a-bf05-fd6c844a98a8 · outbound

This paper cites An argument for basic emotions.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions An argument for basic emotions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.299794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.299794Z digest=sha256:f913344ca1c3dfe88655934dec339cb52692b252e0b9673827848a32e5cb0d98

Observation e43912a0-793e-476a-ab88-55e0bc60b856 · outbound

This paper cites A circumplex model of affect.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions A circumplex model of affect

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.304830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.304830Z digest=sha256:0c799170252ced0c37a9843ee31f2dafcafe1295d0638392f28e00857a6aa816

Observation b75452e2-5d5f-492f-b21f-126db4a70919 · outbound

This paper cites OV-MER: Towards Open-Vocabulary Multimodal Emotion Recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions OV-MER: Towards Open-Vocabulary Multimodal Emotion Recognition

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.309698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.309698Z digest=sha256:4d210902e0396b008900b8f85cc43fdec9f0d46c0bef332d94c654aadd9d9fc4

Observation 8450f0b3-3ee7-4861-a546-d95f9703d13b · outbound

This paper cites GPT-4o System Card.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions GPT-4o System Card

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.314585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.314585Z digest=sha256:812c5959cdaf84feb9e15e7e8441b6b64fe27206488ecf4bc690f81ef28dd25f

Observation 3f4eb40c-0771-4fde-b84e-45e9ef730be2 · outbound

This paper cites Self-report captures 27 distinct categories of emotion bridged by continuous gradients.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Self-report captures 27 distinct categories of emotion bridged by continuous gradients

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.320211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.320211Z digest=sha256:381ac91a5370ec6690f63cc599ba688bbc08d51d2f1387bb6334a7b157249127

Observation 8e58cf6f-20f1-4b0b-bed8-4ba023332c5e · outbound

This paper cites The language of emotion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions The language of emotion

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.324876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.324876Z digest=sha256:b71263cd30e5b3253d76a3603c815f5640f2f54aa4292b6d5af90a741152a7ac

Observation 474ccc17-737c-4a19-b8e1-b08742ecb011 · outbound

This paper cites The role of language in emotion: Predictions from psychological constructionism.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions The role of language in emotion: Predictions from psychological constructionism

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.190794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.329266Z digest=sha256:fe30078f8af0d07ac2df6783aa0e43c2321d0c3f2ee679c016637e207e2c541a

Observation 613824fc-1400-4cf0-ae0d-3f896e0f5bce · outbound

This paper cites Describe your facial expressions by linking image encoders and large language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Describe your facial expressions by linking image encoders and large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.174981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.333575Z digest=sha256:c1c6593ca9595a914037f719386839ad424770f0ff3faa28899e0f1a8d675b25

Observation 6747419b-2eb9-4c31-8d2e-fbce8c61250f · outbound

This paper cites Facial affective behavior analysis with instruction tuning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial affective behavior analysis with instruction tuning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.158768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.337588Z digest=sha256:b75bea3085a17c894c42d22425ff807fbc129b32808fcdeceb5345fecf7550a7

Observation 00e97ee9-0119-45ec-8c4d-9cfcaf6414f9 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning transferable visual models from natural language supervision

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.342458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.342458Z digest=sha256:92b8a2ab71de454e1a85994daa46295e22219c34c38ee2f3c1447e062f2971af

Observation 27759557-c862-47df-b6a3-0975c4adcc68 · outbound

This paper cites Emoclip: A vision-language method for zero-shot video facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Emoclip: A vision-language method for zero-shot video facial expression recognition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.132614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.346762Z digest=sha256:52c4e0326017055fa6e999eeb276d19e3d0dea0e5f046fb6c1580b92c954e452

Observation 63b2af78-2d5c-48e6-a36e-d6e31ff541ff · outbound

This paper cites Flip-80m: 80 million visual-linguistic pairs for facial language-image pre-training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Flip-80m: 80 million visual-linguistic pairs for facial language-image pre-training

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.117499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.350761Z digest=sha256:c54e290e15c79476deb807fc69420bbffa28e4cb0368edc88d1f7ddbbd9a15db

Observation af6ee1a3-6628-434e-8d7d-7bacdc929a0f · outbound

This paper cites Enhancing zero-shot facial expression recognition by llm knowledge transfer.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Enhancing zero-shot facial expression recognition by llm knowledge transfer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.102546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.355611Z digest=sha256:30e678eff0fba88b231d456c599b16da5c9c80c597d6945fea490a6827c3548a

Observation 796b78a5-1e88-4291-95cf-94103d0cd1ba · outbound

This paper cites Facexbench: Evaluating multimodal llms on face understanding.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facexbench: Evaluating multimodal llms on face understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.360923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.360923Z digest=sha256:2bd3af9196361b3e84a1d73f3e69641754957e9d20a30cd2066ee44448c41394

Observation 101f4a02-2524-4b60-84a9-2fee07ea7bc7 · outbound

This paper cites Face-human-bench: A comprehensive benchmark of face and human understanding for multi-modal assistants.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Face-human-bench: A comprehensive benchmark of face and human understanding for multi-modal assistants

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.365402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.365402Z digest=sha256:a2cc00334923f2ce5bb25a4ecc794d121e31e3df2c13e2d786dcaee42eb6b3fc

Observation 9093d8a4-4c21-4974-bc0d-27f203cb14d5 · outbound

This paper cites Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.370480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.370480Z digest=sha256:f862d9419b31fe14d4c54447eb8eac46e33f4d05f3d6203550a15f26be11287a

Observation 3b546bfb-8e31-4df9-9027-375b17c4b38a · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.375535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.375535Z digest=sha256:6509a593f42573dbb6f47eba19e6ab3a607b419232931ce6752872a9cb4ac2f6

Observation 0d4882c9-0f73-42ab-ba71-9383e160156d · outbound

This paper cites Occlusion aware facial expression recognition using cnn with attention mechanism.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Occlusion aware facial expression recognition using cnn with attention mechanism

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.075874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.385225Z digest=sha256:e0055395ff9dd3f31cb54b9b075976a0508ea74d9e71e03384e9ef303ff02588

Observation c26d7ed2-1a9b-4b0f-beab-c05733330d07 · outbound

This paper cites Region attention networks for pose and occlusion robust facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Region attention networks for pose and occlusion robust facial expression recognition

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.061244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.390540Z digest=sha256:2e83192f748e0a3ea1a50282a4d40a9d011c2342af6ca8dfaf18a418863499ba

Observation 430b144b-c4bd-4e05-800f-4dd2e61dbb53 · outbound

This paper cites Learning deep global multi-scale and local attention features for facial expression recognition in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning deep global multi-scale and local attention features for facial expression recognition in the wild

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.044843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.395215Z digest=sha256:aca0965e44118a11fb22eae9f9c8e3506cb03345d6f5f3cfe6bf6f8b16d7326c

Observation 7e04cd3e-a95e-4130-a8b5-01fc21efedf3 · outbound

This paper cites Robust lightweight facial expression recognition network with label distribution training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Robust lightweight facial expression recognition network with label distribution training

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.029929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.399715Z digest=sha256:18aa910565a9877361940214493d770306af80c876b3a869ca348d9233b957f9

Observation 80dd1b37-590b-4991-b264-f22b09697a1a · outbound

This paper cites Facial expression recognition with visual transformers and attentional selective fusion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial expression recognition with visual transformers and attentional selective fusion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:29.014568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.407722Z digest=sha256:82b78a0f5bb288e98944603d5a93137d84d8fee08110317257bb487a80efacdc

Observation c1ea6df3-43d5-425c-a3df-94839e7e39de · outbound

This paper cites Transfer: Learning relation-aware facial expression representations with transformers.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Transfer: Learning relation-aware facial expression representations with transformers

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.998550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.411851Z digest=sha256:903cafbf20a0403c79cc70249037fa5397c0962869e076f054b3a4afdd07f3b2

Observation 4dbb4180-0854-43b3-a52d-afac2bfb917e · outbound

This paper cites Poster: A pyramid cross-fusion transformer network for facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Poster: A pyramid cross-fusion transformer network for facial expression recognition

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.979206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.415946Z digest=sha256:c16120ee96dcc0ed9d0f47d6ae0f17f1ba9c763e4087e6263f292db2685b8d1b

Observation 334c6763-05d5-4a4d-b5c1-ded126b9604a · outbound

This paper cites Svfap: Self-supervised video facial affect perceiver.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Svfap: Self-supervised video facial affect perceiver

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.963118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.420289Z digest=sha256:82e8d7676995d006939d6cd3d324359310873810f1e46bac63869d4e78a0993f

Observation b831ece9-cd9f-4e18-9f18-1e7875b8a1ed · outbound

This paper cites Poster++: A simpler and stronger facial expression recognition network.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Poster++: A simpler and stronger facial expression recognition network

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.943246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.424789Z digest=sha256:b5579d2893afd1c19d69ba03146866fcf2577b5d033acba3f17cbf4fc40d9637

Observation 5d933e14-0c5a-4771-8e24-f997f8411e73 · outbound

This paper cites Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.927537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.431081Z digest=sha256:e3cacc2ef61c3971129d867734cde3ddd0aed05fcb019d8636b16fd9be654c05

Observation d4b4681a-434e-4c5f-89b7-9ff5626202a1 · outbound

This paper cites Suppressing uncertainties for large-scale facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Suppressing uncertainties for large-scale facial expression recognition

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.910444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.435495Z digest=sha256:c769ff996b30cc5d9a5b1c16adeea3c1e0056b5f133cb6093778cac285c1e43c

Observation d9df4532-1db3-4df1-9c92-b1a861982c28 · outbound

This paper cites Relative uncertainty learning for facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Relative uncertainty learning for facial expression recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.892265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.439812Z digest=sha256:2f96bca8d38e35c4c47135ffbd8b5c7c75e378a54df076d221da1208ff3ef63b

Observation ec11e3d6-aea9-44c9-9e48-c902aa3a5b2a · outbound

This paper cites Learning emotion representations from verbal and nonverbal communication.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning emotion representations from verbal and nonverbal communication

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.874331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.444071Z digest=sha256:03f5202ba5dd064685a8d1172fca5e95851cffae81606d1e9fb7872b5d87bcfd

Observation 668b0632-e627-4b26-ac6b-6ece05f3dba8 · outbound

This paper cites Collecting large, richly annotated facial-expression databases from movies.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Collecting large, richly annotated facial-expression databases from movies

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.857788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.448368Z digest=sha256:f6a9a50671f11e28f62f899240f13d0d18f72295c9a72a0a61f0363e60c95b45

Observation 53205689-aba6-4690-b331-ea5483dfff10 · outbound

This paper cites Training deep networks for facial expression recognition with crowd-sourced label distribution.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Training deep networks for facial expression recognition with crowd-sourced label distribution

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.842244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.453976Z digest=sha256:6e7be59faec79fe70dc0d95f35aa9cfc617639a377b41cbf548e7ada5d28d57a

Observation d9730ec3-50fe-4659-b3d4-074cc7ab87f4 · outbound

This paper cites Affectnet: A database for facial expression, valence, and arousal computing in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Affectnet: A database for facial expression, valence, and arousal computing in the wild

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.825975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.458964Z digest=sha256:83792ff895e682c33b00b4eb503ff591f7fd4012cda71b93d7ebcd6a9d9419d4

Observation 2b92e18b-5709-42bd-8711-922081a8989c · outbound

This paper cites Dfew: A large-scale database for recognizing dynamic facial expressions in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Dfew: A large-scale database for recognizing dynamic facial expressions in the wild

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.810550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.463896Z digest=sha256:3d0ee0c8bf37d10592ed46f1381499fe7a1dfbca8d60c4cd1c747793fe071650

Observation e5d8ca84-2fa9-4848-bfc6-bfea2b3c7214 · outbound

This paper cites Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.794142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.468537Z digest=sha256:de702c840e5698d90e9b934e2eea34a1173b614ac04eb26ec795dadaebddd219

Observation 859685ba-3115-451b-8b6d-2e6b1be7dad7 · outbound

This paper cites Ferv39k: A large-scale multi-scene dataset for facial expression recognition in videos.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Ferv39k: A large-scale multi-scene dataset for facial expression recognition in videos

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.777832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.472829Z digest=sha256:614fa86a075199a7a54360fa1aa98bb333d8793e0db4522e9df25835820b6445

Observation a6b0f5c3-de36-4929-8dee-e00c386c8737 · outbound

This paper cites Aff-Wild2: Extending the Aff-Wild Database for Affect Recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Aff-Wild2: Extending the Aff-Wild Database for Affect Recognition

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.477365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.477365Z digest=sha256:efb55d75049b200b96dc99b76e0673b580d06573215d3cf5594c785d05183db1

Observation 45a3adf5-cf83-46dc-adcf-6d8fe902d8c9 · outbound

This paper cites Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.760052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.482389Z digest=sha256:a7f0cb20074ddce9d071b2170f01f76a4492214d50511a14a431d71dc6b13a50

Observation 68a872fe-c5e7-4beb-b4dc-a0bd6f360814 · outbound

This paper cites Compound facial expressions of emotion.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Compound facial expressions of emotion

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.742804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.486532Z digest=sha256:aa99de439a08c9a55e3a625651119ff38014daf0104957146a2acfee103c5d13

Observation bfcf1974-de71-402c-b1eb-b7261af5129e · outbound

This paper cites Configural information in facial expression perception.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Configural information in facial expression perception

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.727222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.491038Z digest=sha256:ba0552084fcbe4903af6abcdd8ccc4b31ad785478b907f2906de8912543baf92

Observation 9c4b2c25-bfde-4039-a2a1-ca4eb145d015 · outbound

This paper cites Parts and wholes in expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Parts and wholes in expression recognition

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.712447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.496626Z digest=sha256:3faa020fee1c40e24338f115ae291acd79684aed5f87f1a554a73599778d29ce

Observation 736555e8-e085-4135-9928-40a153ecaa55 · outbound

This paper cites Mixed emotions: Holistic and analytic perception of facial expressions.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Mixed emotions: Holistic and analytic perception of facial expressions

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.696994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.501743Z digest=sha256:abae450f1d03cd90e39f5a35e8d637f3053266b6387ebaa8962d0c5d114f2d76

Observation c168dbc8-1d94-4a55-baaf-e90bf4d73d38 · outbound

This paper cites The role of facial movements in emotion recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions The role of facial movements in emotion recognition

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.680641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.506388Z digest=sha256:4a031157545bbbc4f93f2d3042fbe2a065c3f4a991d5758dcc6edaeef5fd7bf5

Observation cc17857f-4354-4c3a-afae-b0d9837b3146 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.510794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.510794Z digest=sha256:feaf7f33bbe6b1b35c7f7c7b8174956062752851629c1e87c22ef77e9ff11422

Observation 45b1d21e-5166-4104-b63c-945dd671881e · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Scaling up visual and vision-language representation learning with noisy text supervision

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.515559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.515559Z digest=sha256:62eb5a67a8a70634ad059f4b82e56f87a5b76e218d6cb70cff3f4d1d198a4b96

Observation 543b6e97-02fa-484a-8e72-b90f188ba575 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Reproducible scaling laws for contrastive language-image learning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.519813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.519813Z digest=sha256:188372278536d9d40bae7611a0c8f8970835809d46cef05a7ec9ab9d09578396

Observation 16766737-fb2d-4040-aa01-201cb94a8304 · outbound

This paper cites Sigmoid loss for language image pre-training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Sigmoid loss for language image pre-training

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.524336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.524336Z digest=sha256:f998d083cd3606826affef005ec1991a520ce2cfe22c01f3bb17e2b319232981

Observation a58c779b-930f-4c26-a463-bf3d2c91a56a · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.528659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.528659Z digest=sha256:a98ec4c727af6dc542819f88339e14a444dec996390376f5c1fc8f9e5571caf9

Observation e367f1a0-e1a6-41ea-b20d-c15f86c97ebf · outbound

This paper cites Demystifying clip data.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Demystifying clip data

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.621180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.534253Z digest=sha256:ff88617cdb0564ae6494802658f24d52ca23dd3723d4c89ccf3609cfa29cb260

Observation 296cdcf5-98d0-4dd3-a98d-6a1461eba2ee · outbound

This paper cites Dreamlip: Language-image pre-training with long captions.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Dreamlip: Language-image pre-training with long captions

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.605700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.539781Z digest=sha256:753a71879f13cc2dfc9b9a9c56736bd9ad308d1a2536d01c4921442bafd8302c

Observation 172bed96-dc4c-4ace-b307-44b8a6db6d9e · outbound

This paper cites Modeling Caption Diversity in Contrastive Vision-Language Pretraining.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Modeling Caption Diversity in Contrastive Vision-Language Pretraining

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.544588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.544588Z digest=sha256:ceab434b5cd8f74a4de2c5915248f471626ba78c2d611f10c6b27560935ecd3d

Observation 96e16bd9-f574-4353-8097-d51cf7a79905 · outbound

This paper cites Improving fine-grained understand- ing in image-text pre-training.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Improving fine-grained understand- ing in image-text pre-training

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.590425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.549311Z digest=sha256:bc6470299b5950ca26c3d12aea14a6151f47cb714bcc2422c8df94f5805a433b

Observation 2a162f3a-c265-4461-bf56-c969dce59c59 · outbound

This paper cites General facial representation learning in a visual-linguistic manner.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions General facial representation learning in a visual-linguistic manner

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.554743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.554743Z digest=sha256:57a5d56aab46f630b6c78ccb721337c2d899ce0814d7762f11d8cddf5ccbe493

Observation d1da1679-01f1-4495-bf09-5f9664c2e4ca · outbound

This paper cites GPT-4 Technical Report.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions GPT-4 Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.559162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.559162Z digest=sha256:53cae2d66a52a1ec369f55ffb813c1c23e3c183e5dc8f0f49036ef3e668bc6b2

Observation 62611581-b192-4cca-832b-9a3429ff96ba · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.564018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.564018Z digest=sha256:2bf310ab8c25554768b8056cefc10a2a3f3d2e5cb433b89b36300c7d49212c03

Observation 18510aaf-16b3-425d-8e0a-92146c5d651d · outbound

This paper cites Minigpt-4: Enhancing vision-language understanding with advanced large language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Minigpt-4: Enhancing vision-language understanding with advanced large language models

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.553765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.568575Z digest=sha256:e7528764b24ece0e73fba5701bf18dad9d2ac2b59eca71fca368339d6cfd916b

Observation a7fe66eb-2834-4485-abe6-a6200aecf3bf · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.573362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.573362Z digest=sha256:8f52e1a29ed3b8dc5f9f5a1f16ad00ed2c77082cb2a5030a63947759f45c14bc

Observation 0846c629-f9e1-45ee-94ff-fbfa015973c2 · outbound

This paper cites Qwen Technical Report.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Qwen Technical Report

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.577897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.577897Z digest=sha256:09269ef24c76bb87bb596152743c6a8f7504e0f63fdd72c1872a3376cbe0cb00

Observation b32689e9-eb62-4a25-9e74-442c05fcc889 · outbound

This paper cites Expllm: Towards chain of thought for facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Expllm: Towards chain of thought for facial expression recognition

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.527127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.583046Z digest=sha256:d4cfc0d4c6e5f3e94f6c5aae6159275b32102c05ccf1a8b0f1e8b1dfe98c3d4f

Observation 91706ae5-009f-4165-9389-758e3f05e0f2 · outbound

This paper cites EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions EMO-LLaMA: Enhancing Facial Emotion Understanding with Instruction Tuning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.587777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.587777Z digest=sha256:3b7339f3d621f6d212ebe42e3f76b9a31547490ce26e1335f54c216d3f52bf01

Observation ad3f9e85-5b32-4f95-bb8f-4e6f97606552 · outbound

This paper cites Emotion-llama: Multimodal emotion recognition and reasoning with instruction tuning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Emotion-llama: Multimodal emotion recognition and reasoning with instruction tuning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.510694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.592827Z digest=sha256:fe9c8fc8e5a0570b593a696d9f5237cfe4e7be2025654ba1d4a5c06592e4c74a

Observation 9ff15cd6-87af-4345-9f2f-47c7635f8d50 · outbound

This paper cites AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions AffectGPT: A New Dataset, Model, and Benchmark for Emotion Understanding with Multimodal Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.597574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.597574Z digest=sha256:1485c3b6af89acb5fbf8c0db52e9f876a30c10ae51590a38a44a74bc504eeec8

Observation dcd2a90e-57dd-493c-898a-6e66e0d1b598 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.602296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.602296Z digest=sha256:c7247255dd96fbcbe1f7efeeadd6f74b342625fbc1b1e74df91295bb07cd95c9

Observation 372f1b4f-0d2c-4d48-9411-8e24d1f8fb4e · outbound

This paper cites Generative adversarial network for text-to-face synthesis and manipulation with pretrained bert model.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Generative adversarial network for text-to-face synthesis and manipulation with pretrained bert model

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.493528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.607503Z digest=sha256:2c461c3995523e2ebf434c204f4f3eff44a6413248fb3ee82a52ff8345e6434f

Observation d620aab3-8df8-4240-bf39-ae514f932d14 · outbound

This paper cites Tedigan: Text-guided diverse face image generation and manipulation.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Tedigan: Text-guided diverse face image generation and manipulation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.611815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.611815Z digest=sha256:4d513bc5c6a2d3046fc1a3c1dac2529d79adb25b3ca9642afd4ff8bbada1eac1

Observation bd60d693-e329-40fb-874e-ffc93d0969ba · outbound

This paper cites Talk-to-edit: Fine-grained facial editing via dialog.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Talk-to-edit: Fine-grained facial editing via dialog

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.463190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.616335Z digest=sha256:0fcab1b619a0e68e26062919ac02dc63a0476d37b6423cdbabb48eb36ce578b3

Observation c1535cd7-1311-4ab3-a765-abb76371da81 · outbound

This paper cites 15M Multimodal Facial Image-Text Dataset.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions 15M Multimodal Facial Image-Text Dataset

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.621712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.621712Z digest=sha256:9123f41992642de57ef032f3988307c401b65c830a7e4a42ff6a243cb5ba784c

Observation 100cbd00-c651-4570-a040-0f104650aa21 · outbound

This paper cites Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.626612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.626612Z digest=sha256:5e69c3b5346204e00e68fbc3f9abc12f00f3bd05baba2b2c4a6d751477ad1609

Observation b5516e37-14cc-4f24-95d8-7281e94524eb · outbound

This paper cites Recognizing emotion from facial expressions: psychological and neurological mechanisms.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Recognizing emotion from facial expressions: psychological and neurological mechanisms

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.447457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.632099Z digest=sha256:fe44ed87789e6bf78319a460727edd9430594ad5e350bf656628cc1749175846

Observation c4942f20-be8d-43ad-9381-17fd57fdd3f2 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.636802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.636802Z digest=sha256:9a972d01cc1dec9439f1533e7ce6bdbded99b2d03f386b13beb26735c5c2474a

Observation 49e8119b-58e9-439e-9a72-0efcf4c17de3 · outbound

This paper cites Attention is all you need.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Attention is all you need

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.641491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.641491Z digest=sha256:a5690e4c7ec20998cf526780a97791b661cc10ca0300ae34bd7e5d9d5d2abd7a

Observation 316b960e-eb7b-47f3-b8bd-6f6fc7c5b4d3 · outbound

This paper cites Facial action coding system.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Facial action coding system

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.420990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.646353Z digest=sha256:840919dafd05e099239c73c1a85ef2fb9977a8459f97beb861cbfb41567a77ed

Observation 00c31e87-2361-43ec-9cb8-440841c6c4aa · outbound

This paper cites Softclip: Softer cross-modal alignment makes clip stronger.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Softclip: Softer cross-modal alignment makes clip stronger

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.404940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.651724Z digest=sha256:0a071874466b3229076be22a8159424998432ef38660fdc737c316cd228da119

Observation c92ff191-88ce-47ea-a195-157a86b31f93 · outbound

This paper cites Cwcl: Cross-modal transfer with continuously weighted contrastive loss.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Cwcl: Cross-modal transfer with continuously weighted contrastive loss

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.388027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.658642Z digest=sha256:90fa75b50ba2d9d669f475770100bde406c06d03596947e90e54424c52fe88da

Observation c7333be1-82c7-4c11-ab1e-5c7e57932a2d · outbound

This paper cites Generalizable facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Generalizable facial expression recognition

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.371188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.664046Z digest=sha256:a078eac660494650dedabb2bad7ca68893dc874763eeb585358bf59b3c137eca

Observation c6b01c4a-bc33-4570-bd7d-25213998ac48 · outbound

This paper cites Flava: A foundational language and vision alignment model.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Flava: A foundational language and vision alignment model

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.351256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.669757Z digest=sha256:b7a267aa9ec67173f1e50eebd01bdaefc35d56ab62324117c4b890844d1f76f2

Observation 37e65246-652e-4bcb-92b5-e08e324a1669 · outbound

This paper cites Face-MLLM: A Large Face Perception Model.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Face-MLLM: A Large Face Perception Model

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.675297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.675297Z digest=sha256:166821af34bec1aceb3cda1bda8b0b2c14f4916813be3480e592301df3b5c4b8

Observation 38ee7e1c-45b8-44fe-a6f2-2dbb1596612e · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:27.680305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:27.680305Z digest=sha256:73a44eaa23154531964980e28d66e4924d1c1188f7cb76639989eff4a9a24f8c

Observation 4d20077b-717a-4e68-84af-bfacd77fd4f5 · outbound

This paper cites Learn from all: Erasing attention consistency for noisy label facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learn from all: Erasing attention consistency for noisy label facial expression recognition

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.333222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.686607Z digest=sha256:81520eb51a9e6597211d3e6056eecf35a1f0dddd44bf3d66736c5ddf1a3a92df

Observation 0c7ce87b-05cb-44d0-a49c-09dfc5bd7168 · outbound

This paper cites Latent-ofer: Detect, mask, and reconstruct with latent vectors for occluded facial expression recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Latent-ofer: Detect, mask, and reconstruct with latent vectors for occluded facial expression recognition

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.314024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.692832Z digest=sha256:f32c48b314e65db8d49be35653a152c894287b9aaa6d4868dcc876bbf7b6d7ca

Observation 0ddd7f09-5bdb-4044-92eb-3384ca30156a · outbound

This paper cites From static to dynamic: Adapting landmark-aware image models for facial expression recognition in videos.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions From static to dynamic: Adapting landmark-aware image models for facial expression recognition in videos

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.295881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.698844Z digest=sha256:51357bff5c0076ea7139612635e31b53dec2aa115f558cd5168639a3de4f7aee

Observation 84411ab7-998a-4dea-b93e-156964029dea · outbound

This paper cites Videoclip: Contrastive pre-training for zero-shot video-text understanding.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Videoclip: Contrastive pre-training for zero-shot video-text understanding

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.277285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.705145Z digest=sha256:e15b6886312f14e722b6155031ed2a20aecdd19269dbb469ad43e3fac1a48f27

Observation bb9fa76a-8359-49e1-a6b9-59488c37731a · outbound

This paper cites Expanding language-image pretrained models for general video recognition.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Expanding language-image pretrained models for general video recognition

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.261598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.709988Z digest=sha256:79422d729f8d2b2ed9c49d79666e926a824a42ad27106a3c7b4b8c761a09cb26

Observation 30c11dac-bc73-4cf0-8cae-94b246a1bbc4 · outbound

This paper cites Learning to prompt for vision-language models.

Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions Learning to prompt for vision-language models

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:07:28.243764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T13:07:27.714660Z digest=sha256:fc3dbb5472c977b3f858bbd5daf0bae88f6774a5f2ae718c74b5d5436d9e1291

Pith citing papers

No inbound Pith citation observations are available.