Pith. sign in

Paper Citation Record · LEDGER

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation

As of 19 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2505.18168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18168 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:44:29.296392Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd585619-4881-4431-a9bd-fe730379ea18 · outbound

This paper cites A comprehensive review of facial expression recognition techniques.Multimedia Systems, 29(1):73–103, 2023.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A comprehensive review of facial expression recognition techniques.Multimedia Systems, 29(1):73–103, 2023

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.448939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.894066Z digest=sha256:b747b689076f399fe44360efc63a7648a52c845ccca84d7b17f898932c9385b3

Observation f4c81d0c-efae-426b-84af-0d4c07a6e459 · outbound

This paper cites Llama 3 model card.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Llama 3 model card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.900814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.900814Z digest=sha256:f52e30b70f50f5d7d315ee5c20d1d807b1bf0400c77a0bf49ffe184b6a4c0ca4

Observation f9b005ec-5505-41f0-9761-d91329bffaf0 · outbound

This paper cites Claude-3.5.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Claude-3.5

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.909847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.909847Z digest=sha256:465c0bfa93eab94b22110aa7661cd66be4959e67a8a3ee1cc1529ce298de5708

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.915865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.915865Z digest=sha256:95115e8de376c491195f87faaac2e950b0643ff1c5ee38924d9c8cb4c4291c14

Observation efa85399-bc12-401f-8eb7-10c7d6329949 · outbound

This paper cites Video-based facial micro-expression analysis: A survey of datasets, features and algorithms.IEEE transactions on pattern analysis and machine intelligence, 44(9):5826–5846, 2021.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Video-based facial micro-expression analysis: A survey of datasets, features and algorithms.IEEE transactions on pattern analysis and machine intelligence, 44(9):5826–5846, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.393790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.923638Z digest=sha256:cf54ef5a00185d6758089c879f92ed21883635d965d85cd3dfdd54fb4772b0fc

Observation f1a0ea46-782d-486d-9fd2-3cc0bb322616 · outbound

This paper cites Knowledge-driven self-supervised representa- tion learning for facial action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge-driven self-supervised representa- tion learning for facial action unit recognition

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.373312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.932420Z digest=sha256:f954430cc39dc1aa771c258a234047857245859f0244f536fa3b96416b2fd9b8

Observation 64dff31f-2c95-413c-bc29-a408f7a74e10 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.941071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.941071Z digest=sha256:3acd4305b83a758e628083392ed03931638c8456fa256563cf946db131382904

Observation 6df65fbd-105c-42ed-bcb9-d5ad5940160a · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.948070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.948070Z digest=sha256:11ac86a296b69a91f8487f256833b26a598bdc62ea5f500aa1a5985dda34cadb

Observation c113b1d2-cb90-4053-9e67-7f9642cd5069 · outbound

This paper cites Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.953395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.953395Z digest=sha256:6fc4a52fd3c0fafdd152252066357692edbae4d595baef63c1d78616df32fe06

Observation 6480a781-f225-4d80-bd6a-e0962068275f · outbound

This paper cites Knowledge augmented deep neural networks for joint facial expression and action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge augmented deep neural networks for joint facial expression and action unit recognition

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.357090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.959369Z digest=sha256:ac2ef7f30076dc24e37d252376c97106af982dab150090e25edd9283812e74e6

Observation 0b40f4f7-85c6-414f-93d8-9a224ae0e826 · outbound

This paper cites Knowledge augmented deep neural networks for joint facial expression and action unit recognition.Advances in Neural Information Processing Systems, 33:14338–14349, 2020.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge augmented deep neural networks for joint facial expression and action unit recognition.Advances in Neural Information Processing Systems, 33:14338–14349, 2020

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.338683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.965617Z digest=sha256:3d5b10aa6444fa9111d5e8d1640375f741c777ede3a9f078bde50559ce75e8ca

Observation 70b427b1-5d35-4377-8650-259d7ecd28e4 · outbound

This paper cites Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.319853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.971272Z digest=sha256:80152bd3791ec976dc0a35092350805c6f29a39ea264fb8de96e0780b7791472

Observation ad635b35-ec6a-4bfd-963f-c24dba3eaed2 · outbound

This paper cites Consulting Psychologists Press, 1978.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Consulting Psychologists Press, 1978

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.302560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.976805Z digest=sha256:fac30bc71d117185aa714f693be44f8880f0ffddb5e776f4bb8e6c3f7d2c06ee

Observation e0164aa9-a2d8-4f3b-8ba0-e1a5767d1518 · outbound

This paper cites Universals and cultural differences in the judgments of facial expressions of emotion.Journal of personality and social psychology, 53(4):712, 1987.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Universals and cultural differences in the judgments of facial expressions of emotion.Journal of personality and social psychology, 53(4):712, 1987

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.284906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.981797Z digest=sha256:952bfbfc026ca3bf59f02ddf15b3ab34d46c2c1b55bc586ddf94d3345bbc5cc9

Observation 75196822-5c1b-464b-9ded-20435cad4d14 · outbound

This paper cites Oxford University Press, USA, 1997.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Oxford University Press, USA, 1997

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.267463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.988199Z digest=sha256:a2afb8c952f4a4308435fffb543a6792156d798d0f2552eea74e0ebe8685f7fe

Observation 14a49383-e8f3-456d-8012-77ed47e06e97 · outbound

This paper cites Selfme: Self-supervised motion learning for micro-expression recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Selfme: Self-supervised motion learning for micro-expression recognition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.249255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:28.995942Z digest=sha256:e7f8462cf227ec8904d4b59a60a0dd67a48b0cdbd81f039c34c2df534fba6074

Observation 3da9ee51-14c9-4a74-bb20-d6377e3075db · outbound

This paper cites StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.001846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.001846Z digest=sha256:d544ab3fd53d8f4829fd46ddc9bea12b342a5d4501b040dc9373f42611f47438

Observation 3a11e821-1bf5-4677-beb8-bdae00a73d9f · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.225745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.007507Z digest=sha256:aad73a9f867e9ed6e563fafdd56ec945b8c5b73db7d6e4ba01369f14ebb5c7cd

Observation 57c4d9c6-abde-4ca6-a54f-8f3c02e4eb99 · outbound

This paper cites Disentangling identity and pose for facial ex- pression recognition.IEEE Transactions on Affective Computing, 13(4):1868–1878, 2022.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Disentangling identity and pose for facial ex- pression recognition.IEEE Transactions on Affective Computing, 13(4):1868–1878, 2022

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.205050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.013511Z digest=sha256:c9c702b9c03ea2978aada636a9ed283a6413612bd781eb7bb073c49bb1cbb588

Observation 745b3899-9528-4f52-b83f-f36bb18e897b · outbound

This paper cites Scaling Laws for Neural Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Scaling Laws for Neural Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.020890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.020890Z digest=sha256:06a87c281ef39279a18beea0378645d57cadbf0b7db96c3697c16a3e3b2465ce

Observation 4dbe5b7c-5255-4780-858c-4b1e63e723ed · outbound

This paper cites What uncertainties do we need in bayesian deep learning for computer vision?Advances in neural information processing systems, 30, 2017.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation What uncertainties do we need in bayesian deep learning for computer vision?Advances in neural information processing systems, 30, 2017

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.027335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.027335Z digest=sha256:72d25d896d75973e778da8dd1273f272b8e48fe438d880d25109ab505bfa45ad

Observation 14e3cbe0-f036-49c6-9812-b2f827fa4ebe · outbound

This paper cites Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.032686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.032686Z digest=sha256:266ea4df1ac5fafcf105a36e1485761c13aebf7503f5183b70bc36b4ab388abf

Observation e131d91b-6d0a-42ba-9509-7a97a75e051a · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaVA-OneVision: Easy Visual Task Transfer

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.038401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.038401Z digest=sha256:a287de5b1134ddfdd2722d7c80394b03d5637da1c63e23766691d71b1eb66f3e

Observation b9df44d5-067c-460b-a979-2be5b9b89ed7 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.051757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.051757Z digest=sha256:4e490eafe09591ababb49ac1ac42a62e5c9417f199e94477d4a44946f5df6044

Observation f7e0f455-66ff-43f4-b495-efc396b13ea6 · outbound

This paper cites Deep facial expression recognition: A survey.IEEE transactions on affective computing, 13(3):1195–1215, 2020.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Deep facial expression recognition: A survey.IEEE transactions on affective computing, 13(3):1195–1215, 2020

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.168425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.057474Z digest=sha256:c29dd910943349b6afe665bf9564879f71e4316b1097a93b4034b975d0c88abc

Observation fc6dc9f8-2aa3-4003-a90e-44dd3b59918c · outbound

This paper cites Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.062498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.062498Z digest=sha256:fcdc966d697eac2a1fef149cd0395f827091312753573a3aefa09b065a77054f

Observation 9be963ef-447c-4508-bf1d-b1e0c7f15dce · outbound

This paper cites AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.077312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.077312Z digest=sha256:85bad16f29d93d0f3073436abfe9cee6d1230b2221a1cbb7889d07335df1beb8

Observation 98a81cd0-be0a-462b-ab95-3b99c413f73b · outbound

This paper cites Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.Information Fusion, 108:102367, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.Information Fusion, 108:102367, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.132961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.084190Z digest=sha256:cec2fbc475c745a14970ceb9dfc6cbf2688aab9d9a4e9899de6a42e592b3109f

Observation a712980b-a266-4421-a2fc-2ed05ad21a9f · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.089156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.089156Z digest=sha256:1677b76cc99a0d51385a715cc1816e15021b3113c4cdf39ab3c68e3deeace1b0

Observation dbf13a1f-c0d8-4dbe-b698-d95295c02381 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:44:30.113934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.094388Z digest=sha256:84989662426b65c913c4e51b82999af5870908baf63fb55aae5e3a8389db06f6

Observation 51b2a338-d5b0-4a45-bdc6-eebed7f88f3c · outbound

This paper cites Improved baselines with visual instruction tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Improved baselines with visual instruction tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.099555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.099555Z digest=sha256:fd96af79dc05953abf27e0c94fe137a078b09aba1ea1ed69a303bff87124828c

Observation e3975bf9-e31e-4837-b8cf-bd9194cfc3c3 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.105455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.105455Z digest=sha256:1c45894e17b4becf16c3acaa0a2f893f64fe6c1623b1b9fd77105d8d298ec959

Observation bcefcfd6-9b7b-433c-84e8-a4efc059aea2 · outbound

This paper cites Facial expressions elicit multiplexed perceptions of emotion categories and dimensions.Current Biology, 32(1):200–209, 2022.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Facial expressions elicit multiplexed perceptions of emotion categories and dimensions.Current Biology, 32(1):200–209, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.047620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.113610Z digest=sha256:762cac126f4013adeba022747df2f953cb236ecf45f6763f7bf30665311c18b5

Observation 77c80aab-deea-4ab0-b651-fb41d2a2aedb · outbound

This paper cites Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.121054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.121054Z digest=sha256:fb005fd7f6b00403d3f2f969e359fc1c0d672b0f7c5a4be6695c6d96f6d729fb

Observation 125f7e72-fa0f-4036-9d25-3013e45fa241 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:44:30.018532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.127122Z digest=sha256:231f8945d872cc68c0ccae7e4bd618c72e6395caf3070b921ed455b24a798663

Observation 79a82334-a25c-44e0-b9bd-6e1dac34184c · outbound

This paper cites Disfa: A spontaneous facial action intensity database.IEEE Transactions on Affective Computing, 4(2):151–160, 2013.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Disfa: A spontaneous facial action intensity database.IEEE Transactions on Affective Computing, 4(2):151–160, 2013

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.996354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.136191Z digest=sha256:aed38fc3701593fa4760f99058ff82cd4edab22691087979766e465b0a1a8dd5

Observation 8054be0b-be9f-4196-a4e4-dd3a699d5405 · outbound

This paper cites Affectnet: A database for facial expression, valence, and arousal computing in the wild.IEEE Transactions on Affective Computing, 10(1):18–31, 2017.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Affectnet: A database for facial expression, valence, and arousal computing in the wild.IEEE Transactions on Affective Computing, 10(1):18–31, 2017

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.976436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.141926Z digest=sha256:79513b3d25485acfc8302ddd090ca551951b19f48464fe4461834e5b350dbb06

Observation 979104c4-dc8e-4e2f-90d1-45aeafc68fa1 · outbound

This paper cites Multi-label co- regularization for semi-supervised facial action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Multi-label co- regularization for semi-supervised facial action unit recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.955566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.146858Z digest=sha256:9adfbd21c98ee025b9a0b644d96950db40d62a3e573c264acf21d6aadfaeb838

Observation 5901f188-0ce3-4085-9712-84e3446d8ec2 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.152898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.152898Z digest=sha256:97442b589744221130b667146777f5e396e5dac2b65b0d81ca4c6fe09d6bea59

Observation 2eda21f1-d106-4c47-ab59-be000aac1d17 · outbound

This paper cites Hello gpt-4o.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Hello gpt-4o

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.158753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.158753Z digest=sha256:389e3f9dbdd4a7abdf8f3b1f3b18ed9a02db2afaaf1d26b6d70b23e2d015fc6e

Observation a10c2b5b-932d-419c-b748-35c1eab6f891 · outbound

This paper cites A unified and interpretable emotion representation and expression generation.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A unified and interpretable emotion representation and expression generation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.908899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.165525Z digest=sha256:597e51c8134502b24861a41c072ec06eea5cf6341afb37b38a7852edb33b4783

Observation d5cd58f2-4a73-469d-982c-3bc941549031 · outbound

This paper cites A circumplex model of affect.Journal of personality and social psychology, 39(6):1161, 1980.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A circumplex model of affect.Journal of personality and social psychology, 39(6):1161, 1980

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.178279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.178279Z digest=sha256:9d054bec884ab51de57f6a7025d841fc48e79a89eb24f93601a7e170ea015818

Observation edb27f16-0e2b-4d94-a18e-edc32ab125a2 · outbound

This paper cites Uncertain graph neural networks for facial action unit detection.Proceedings of the AAAI Conference on Artificial Intelligence, 35(7):5993–6001, 2021.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Uncertain graph neural networks for facial action unit detection.Proceedings of the AAAI Conference on Artificial Intelligence, 35(7):5993–6001, 2021

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.875105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.187860Z digest=sha256:393febbfb7da80b325ba57fe95b7a2288a7cc673885745517db1b710f24764e6

Observation a8f47576-8091-4484-9c08-adbeb174d240 · outbound

This paper cites Hybrid message passing with performance-driven structures for facial action unit detection.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Hybrid message passing with performance-driven structures for facial action unit detection

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.854631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.194648Z digest=sha256:9c0160cacddfc6cf28169d99c1bcee9abd2fafda1f69a85f2078bc3a0f8777f5

Observation d3416001-0052-425e-a3a4-23903835f89c · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gemini: A Family of Highly Capable Multimodal Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.199625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.199625Z digest=sha256:febe2bdbb688c06c0a2a3fbfd62c45bf599d7eaecc0ff52daf40817c6846845f

Observation 3fe20492-2509-484f-9a1b-b6ab00c93242 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.207681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.207681Z digest=sha256:2b6b6450620a0b2b2c5bc5649a5420b6836fc1f51c5c0b3fdccb3484a3426c86

Observation 3aae1250-6c06-48f5-b557-3bba229d829b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaMA: Open and Efficient Foundation Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.212990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.212990Z digest=sha256:38c22dccf9abf0c44ecba98e2c62cf62603ed876a1ddc7cf0f00882f03a82326

Observation 0dd3b1f7-89e4-4717-a477-52605470f60e · outbound

This paper cites Rethinking the learning paradigm for dynamic facial expression recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Rethinking the learning paradigm for dynamic facial expression recognition

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.831395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.218238Z digest=sha256:04773264bf3adafc54676eb3fba76f4998fb920bdb9404e3580eef17563980ac

Observation 35934a95-d9e1-42a0-a6dc-1c0ce49d7288 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.223723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.223723Z digest=sha256:cf5f9371c57da70434cea676351f28269eafc687d2b89287d8cad19a847d6338

Observation de36e934-6782-42fa-9aee-d879762dbda5 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.233811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.233811Z digest=sha256:8acd23dcfb85cd54b93a94c315d79bc3dd1dc6e21491a4180495c9dd574f3950

Observation f76c93ae-0f6e-42bc-ae1f-c6ae17b41707 · outbound

This paper cites Emovit: Revolutionizing emotion insights with visual instruction tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Emovit: Revolutionizing emotion insights with visual instruction tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.240611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.240611Z digest=sha256:caae065ad4610842ff43e8f0565a7bd22bb282e315629b5e50bfac1a05649406

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.247106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.247106Z digest=sha256:610c506f20a1abeff06e0d2892e49d62175885f58541dcb60f5d50a1d85d8cbc

Observation c69de7ab-bd21-4e78-969f-d0d6795b13a4 · outbound

This paper cites Robust emotion recognition in context debiasing.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Robust emotion recognition in context debiasing

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.793904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.255486Z digest=sha256:bd6cbd2e0d53bf143a78f3c40ea045bda78774c7488fe2ccfb213d381bf6f8d3

Observation 492de562-dca7-4574-bf19-7a55fa8adf0d · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.262550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.262550Z digest=sha256:2dad19e85ead8dd745d4a5557abe03256c06f9b57d2fff7af922e16a07ef36b0

Observation aebe6239-57eb-4344-ada0-bd843454f20a · outbound

This paper cites Sigmoid loss for language image pre-training.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Sigmoid loss for language image pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.276459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.276459Z digest=sha256:3441f25f74a5dcb7b505fd012c5d34eb308bfc031d89aaa2bea7a106c0b3b216

Observation c66b5d0a-dbb2-4932-9697-f7fb8ebd4415 · outbound

This paper cites A high-resolution spontaneous 3d dynamic facial expression database.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A high-resolution spontaneous 3d dynamic facial expression database

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.761442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.284777Z digest=sha256:1c413c655b7ea30f55069d530c4c0c7cd4404eac2bb4cedb17dd83c661bce90c

Observation 45d9dd06-f5d7-46be-809d-79e895bd6a15 · outbound

This paper cites Bp4d-spontaneous: a high-resolution spontaneous 3d dynamic facial expression database.Image and Vision Computing, 32(10):692–706, 2014.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Bp4d-spontaneous: a high-resolution spontaneous 3d dynamic facial expression database.Image and Vision Computing, 32(10):692–706, 2014

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.290214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.290214Z digest=sha256:319336b45845ffa94468c4d99a1c4410abd5134fb757780739ed6ac495c92b0e

Observation d99c2338-975d-412d-81a8-97d9fad1dd0c · outbound

This paper cites Khfa: Knowledge-driven hierarchical feature alignment framework for subject- invariant facial action unit detection.IEEE Transactions on Instrumentation and Measurement, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Khfa: Knowledge-driven hierarchical feature alignment framework for subject- invariant facial action unit detection.IEEE Transactions on Instrumentation and Measurement, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.730776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T21:44:29.296392Z digest=sha256:ae0a6d0bb90c08a57618338d09d069c7f4adbd101c2ceea66d15e8f102d3454a

Pith citing papers

No inbound Pith citation observations are available.