Pith. sign in

Paper Citation Record · LEDGER

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation

As of 16 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2505.18168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18168 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:44:29.296392Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved33
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd585619-4881-4431-a9bd-fe730379ea18 · outbound

This paper cites A comprehensive review of facial expression recognition techniques.Multimedia Systems, 29(1):73–103, 2023.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A comprehensive review of facial expression recognition techniques.Multimedia Systems, 29(1):73–103, 2023

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.448939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.894066Z digest=sha256:3180735ddee9cda8a5ef80fcbe3f09909448d2c9a5cdeaea89c809422a593dae

Observation f4c81d0c-efae-426b-84af-0d4c07a6e459 · outbound

This paper cites Llama 3 model card.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Llama 3 model card

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.900814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.900814Z digest=sha256:f52e30b70f50f5d7d315ee5c20d1d807b1bf0400c77a0bf49ffe184b6a4c0ca4

Observation f9b005ec-5505-41f0-9761-d91329bffaf0 · outbound

This paper cites Claude-3.5.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Claude-3.5

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.909847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.909847Z digest=sha256:465c0bfa93eab94b22110aa7661cd66be4959e67a8a3ee1cc1529ce298de5708

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.915865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.915865Z digest=sha256:95115e8de376c491195f87faaac2e950b0643ff1c5ee38924d9c8cb4c4291c14

Observation efa85399-bc12-401f-8eb7-10c7d6329949 · outbound

This paper cites Video-based facial micro-expression analysis: A survey of datasets, features and algorithms.IEEE transactions on pattern analysis and machine intelligence, 44(9):5826–5846, 2021.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Video-based facial micro-expression analysis: A survey of datasets, features and algorithms.IEEE transactions on pattern analysis and machine intelligence, 44(9):5826–5846, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.393790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.923638Z digest=sha256:7e7b69fdd1dc1e24748b511955319273f12c72e43bcf8b1e8c435d5591af2c14

Observation f1a0ea46-782d-486d-9fd2-3cc0bb322616 · outbound

This paper cites Knowledge-driven self-supervised representa- tion learning for facial action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge-driven self-supervised representa- tion learning for facial action unit recognition

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.373312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.932420Z digest=sha256:5ef2a139a3b09d5581867ee83b76d5510d356aa50a31f69061d420cb62631bd7

Observation 64dff31f-2c95-413c-bc29-a408f7a74e10 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.941071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.941071Z digest=sha256:8318f5f4ed5ba9093fc1de2a3ea8bc8621fe3b150b9b764f245b37e15fc565fe

Observation 6df65fbd-105c-42ed-bcb9-d5ad5940160a · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.948070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.948070Z digest=sha256:423d6e43825456f62aaa2aa8dcb6cef76628c756b20da49b317a0c57e388bd25

Observation c113b1d2-cb90-4053-9e67-7f9642cd5069 · outbound

This paper cites Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Emotion-LLaMA: Multimodal Emotion Recognition and Reasoning with Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:28.953395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:28.953395Z digest=sha256:6706d89aca82dd52ef169384f0e61b639d1bd20fc5434b4719449deb1221135d

Observation 6480a781-f225-4d80-bd6a-e0962068275f · outbound

This paper cites Knowledge augmented deep neural networks for joint facial expression and action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge augmented deep neural networks for joint facial expression and action unit recognition

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.357090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.959369Z digest=sha256:e01b8718078a55f19536bc710658fecda4a13aaf12677f48d8ddda593664179f

Observation 0b40f4f7-85c6-414f-93d8-9a224ae0e826 · outbound

This paper cites Knowledge augmented deep neural networks for joint facial expression and action unit recognition.Advances in Neural Information Processing Systems, 33:14338–14349, 2020.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Knowledge augmented deep neural networks for joint facial expression and action unit recognition.Advances in Neural Information Processing Systems, 33:14338–14349, 2020

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.338683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.965617Z digest=sha256:e5646326651bc1a26466971d37160f7ef3e035805a85d07fbae1b40cd70e73ef

Observation 70b427b1-5d35-4377-8650-259d7ecd28e4 · outbound

This paper cites Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Decomposition of uncertainty in bayesian deep learning for efficient and risk-sensitive learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.319853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.971272Z digest=sha256:3a3a2bbb7993792e636bf4fa07eadb45d37b44b7e64c012f2c073edef2738f3a

Observation ad635b35-ec6a-4bfd-963f-c24dba3eaed2 · outbound

This paper cites Consulting Psychologists Press, 1978.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Consulting Psychologists Press, 1978

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.302560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.976805Z digest=sha256:f3545be0fdd754e24386df329d5c1b1515d6fcf1eaf4b64828c198387dd051ba

Observation e0164aa9-a2d8-4f3b-8ba0-e1a5767d1518 · outbound

This paper cites Universals and cultural differences in the judgments of facial expressions of emotion.Journal of personality and social psychology, 53(4):712, 1987.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Universals and cultural differences in the judgments of facial expressions of emotion.Journal of personality and social psychology, 53(4):712, 1987

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.284906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.981797Z digest=sha256:ef42d684e2345646de6f9e4b97893911bc4808fd245a55af1c520cebfebb8627

Observation 75196822-5c1b-464b-9ded-20435cad4d14 · outbound

This paper cites Oxford University Press, USA, 1997.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Oxford University Press, USA, 1997

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.267463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.988199Z digest=sha256:3ac88a2694c0d7f38ecbd43a7f2be9d0c8bb7ecb792915130cafe3f3ee4350c2

Observation 14a49383-e8f3-456d-8012-77ed47e06e97 · outbound

This paper cites Selfme: Self-supervised motion learning for micro-expression recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Selfme: Self-supervised motion learning for micro-expression recognition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.249255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:28.995942Z digest=sha256:97696a2c5b26e26434ba354e86702e7532e00db41d07cfd1277c05e4c622f84b

Observation 3da9ee51-14c9-4a74-bb20-d6377e3075db · outbound

This paper cites StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.001846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.001846Z digest=sha256:c5eef4599c0c9230f0fb0f5703c60138bfd8eed39dc19664f8b07eef6a5c1826

Observation 3a11e821-1bf5-4677-beb8-bdae00a73d9f · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.225745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.007507Z digest=sha256:36fd0b7dcc01db6a67c45253aeef34b4ce89b14e659075149715cbe6d86ea3d0

Observation 57c4d9c6-abde-4ca6-a54f-8f3c02e4eb99 · outbound

This paper cites Disentangling identity and pose for facial ex- pression recognition.IEEE Transactions on Affective Computing, 13(4):1868–1878, 2022.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Disentangling identity and pose for facial ex- pression recognition.IEEE Transactions on Affective Computing, 13(4):1868–1878, 2022

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.205050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.013511Z digest=sha256:3504ee88f96579ccc4deb70ca274a82e974b74675b358613ad2cb6f104cc6876

Observation 745b3899-9528-4f52-b83f-f36bb18e897b · outbound

This paper cites Scaling Laws for Neural Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Scaling Laws for Neural Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.020890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.020890Z digest=sha256:06a87c281ef39279a18beea0378645d57cadbf0b7db96c3697c16a3e3b2465ce

Observation 4dbe5b7c-5255-4780-858c-4b1e63e723ed · outbound

This paper cites What uncertainties do we need in bayesian deep learning for computer vision?Advances in neural information processing systems, 30, 2017.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation What uncertainties do we need in bayesian deep learning for computer vision?Advances in neural information processing systems, 30, 2017

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.027335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.027335Z digest=sha256:72d25d896d75973e778da8dd1273f272b8e48fe438d880d25109ab505bfa45ad

Observation 14e3cbe0-f036-49c6-9812-b2f827fa4ebe · outbound

This paper cites Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Expression, Affect, Action Unit Recognition: Aff-Wild2, Multi-Task Learning and ArcFace

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.032686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.032686Z digest=sha256:266ea4df1ac5fafcf105a36e1485761c13aebf7503f5183b70bc36b4ab388abf

Observation e131d91b-6d0a-42ba-9509-7a97a75e051a · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaVA-OneVision: Easy Visual Task Transfer

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.038401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.038401Z digest=sha256:0ba3f845025d96be036a75bed580166c99e443a4663ab08f8e6bb307e85e50c9

Observation b9df44d5-067c-460b-a979-2be5b9b89ed7 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.051757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.051757Z digest=sha256:4e490eafe09591ababb49ac1ac42a62e5c9417f199e94477d4a44946f5df6044

Observation f7e0f455-66ff-43f4-b495-efc396b13ea6 · outbound

This paper cites Deep facial expression recognition: A survey.IEEE transactions on affective computing, 13(3):1195–1215, 2020.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Deep facial expression recognition: A survey.IEEE transactions on affective computing, 13(3):1195–1215, 2020

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.168425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.057474Z digest=sha256:97610a5d5eaefaf9d4a4df84e3a11c7e28067a68d132881d7784a8cfada3a963

Observation fc6dc9f8-2aa3-4003-a90e-44dd3b59918c · outbound

This paper cites Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.062498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.062498Z digest=sha256:fcdc966d697eac2a1fef149cd0395f827091312753573a3aefa09b065a77054f

Observation 9be963ef-447c-4508-bf1d-b1e0c7f15dce · outbound

This paper cites AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation AffectGPT: Dataset and Framework for Explainable Multimodal Emotion Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.077312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.077312Z digest=sha256:c73700e8350206f995e94f049dade63f1cd2d69c7c4e96381146436627ce0344

Observation 98a81cd0-be0a-462b-ab95-3b99c413f73b · outbound

This paper cites Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.Information Fusion, 108:102367, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gpt-4v with emotion: A zero-shot benchmark for generalized emotion recognition.Information Fusion, 108:102367, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.132961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.084190Z digest=sha256:f024a153e4441c0fdedbd1a9f97282c0339f1d9cc8204a607cdfabfc26b81eb4

Observation a712980b-a266-4421-a2fc-2ed05ad21a9f · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.089156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.089156Z digest=sha256:1677b76cc99a0d51385a715cc1816e15021b3113c4cdf39ab3c68e3deeace1b0

Observation dbf13a1f-c0d8-4dbe-b698-d95295c02381 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:44:30.113934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.094388Z digest=sha256:1fd9adb8f680a67eec8598f72d38ade9616a51ddce58f091177e24cf555a56f0

Observation 51b2a338-d5b0-4a45-bdc6-eebed7f88f3c · outbound

This paper cites Improved baselines with visual instruction tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Improved baselines with visual instruction tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.099555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.099555Z digest=sha256:fd96af79dc05953abf27e0c94fe137a078b09aba1ea1ed69a303bff87124828c

Observation e3975bf9-e31e-4837-b8cf-bd9194cfc3c3 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Visual instruction tuning.Advances in neural information processing systems, 36, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.105455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.105455Z digest=sha256:1c45894e17b4becf16c3acaa0a2f893f64fe6c1623b1b9fd77105d8d298ec959

Observation bcefcfd6-9b7b-433c-84e8-a4efc059aea2 · outbound

This paper cites Facial expressions elicit multiplexed perceptions of emotion categories and dimensions.Current Biology, 32(1):200–209, 2022.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Facial expressions elicit multiplexed perceptions of emotion categories and dimensions.Current Biology, 32(1):200–209, 2022

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:30.047620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.113610Z digest=sha256:50e0e07dc53a4c71ec9ec6655f4d7920007bee354bb95f16babc3100e97d1b25

Observation 77c80aab-deea-4ab0-b651-fb41d2a2aedb · outbound

This paper cites Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Mafw: A large-scale, multi-modal, compound affective database for dynamic facial expression recognition in the wild

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.121054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.121054Z digest=sha256:fb005fd7f6b00403d3f2f969e359fc1c0d672b0f7c5a4be6695c6d96f6d729fb

Observation 125f7e72-fa0f-4036-9d25-3013e45fa241 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:44:30.018532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.127122Z digest=sha256:0cb27f91b17ca49f80dff1b4b409d844b6a4c9f88676a24b6066828d547943dd

Observation 79a82334-a25c-44e0-b9bd-6e1dac34184c · outbound

This paper cites Disfa: A spontaneous facial action intensity database.IEEE Transactions on Affective Computing, 4(2):151–160, 2013.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Disfa: A spontaneous facial action intensity database.IEEE Transactions on Affective Computing, 4(2):151–160, 2013

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.996354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.136191Z digest=sha256:828cd53748594e1b02e0b0696ff5e595f4c55d94a745d5c577965c0d46853669

Observation 8054be0b-be9f-4196-a4e4-dd3a699d5405 · outbound

This paper cites Affectnet: A database for facial expression, valence, and arousal computing in the wild.IEEE Transactions on Affective Computing, 10(1):18–31, 2017.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Affectnet: A database for facial expression, valence, and arousal computing in the wild.IEEE Transactions on Affective Computing, 10(1):18–31, 2017

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.976436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.141926Z digest=sha256:636069a8988f1f72cbac47825c0c8d9a76f12f380073c677c87a5ed1ea13c933

Observation 979104c4-dc8e-4e2f-90d1-45aeafc68fa1 · outbound

This paper cites Multi-label co- regularization for semi-supervised facial action unit recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Multi-label co- regularization for semi-supervised facial action unit recognition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.955566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.146858Z digest=sha256:8e3a76498c6439e03e0b842f29b822508bf83f838531253f517a5865040f5928

Observation 5901f188-0ce3-4085-9712-84e3446d8ec2 · outbound

This paper cites an unresolved cited work.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.152898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.152898Z digest=sha256:97442b589744221130b667146777f5e396e5dac2b65b0d81ca4c6fe09d6bea59

Observation 2eda21f1-d106-4c47-ab59-be000aac1d17 · outbound

This paper cites Hello gpt-4o.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Hello gpt-4o

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.158753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.158753Z digest=sha256:389e3f9dbdd4a7abdf8f3b1f3b18ed9a02db2afaaf1d26b6d70b23e2d015fc6e

Observation a10c2b5b-932d-419c-b748-35c1eab6f891 · outbound

This paper cites A unified and interpretable emotion representation and expression generation.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A unified and interpretable emotion representation and expression generation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.908899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.165525Z digest=sha256:8ce3f7fe6f956c7b7b9e4d38a1157a55bee5f96fdf63cd7a81676f6f7bd04b71

Observation d5cd58f2-4a73-469d-982c-3bc941549031 · outbound

This paper cites A circumplex model of affect.Journal of personality and social psychology, 39(6):1161, 1980.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A circumplex model of affect.Journal of personality and social psychology, 39(6):1161, 1980

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.178279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.178279Z digest=sha256:9d054bec884ab51de57f6a7025d841fc48e79a89eb24f93601a7e170ea015818

Observation edb27f16-0e2b-4d94-a18e-edc32ab125a2 · outbound

This paper cites Uncertain graph neural networks for facial action unit detection.Proceedings of the AAAI Conference on Artificial Intelligence, 35(7):5993–6001, 2021.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Uncertain graph neural networks for facial action unit detection.Proceedings of the AAAI Conference on Artificial Intelligence, 35(7):5993–6001, 2021

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.875105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.187860Z digest=sha256:c64afcd0cc7b79731fbfb97b6d407eaab8ad13047167a45a0557f57bcbe8afe3

Observation a8f47576-8091-4484-9c08-adbeb174d240 · outbound

This paper cites Hybrid message passing with performance-driven structures for facial action unit detection.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Hybrid message passing with performance-driven structures for facial action unit detection

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.854631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.194648Z digest=sha256:c3da7cae0f6eb1e4e8257fa83b780cd7591eb735f00456a85d92035ca678378e

Observation d3416001-0052-425e-a3a4-23903835f89c · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gemini: A Family of Highly Capable Multimodal Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.199625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.199625Z digest=sha256:febe2bdbb688c06c0a2a3fbfd62c45bf599d7eaecc0ff52daf40817c6846845f

Observation 3fe20492-2509-484f-9a1b-b6ab00c93242 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.207681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.207681Z digest=sha256:2b6b6450620a0b2b2c5bc5649a5420b6836fc1f51c5c0b3fdccb3484a3426c86

Observation 3aae1250-6c06-48f5-b557-3bba229d829b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation LLaMA: Open and Efficient Foundation Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.212990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.212990Z digest=sha256:38c22dccf9abf0c44ecba98e2c62cf62603ed876a1ddc7cf0f00882f03a82326

Observation 0dd3b1f7-89e4-4717-a477-52605470f60e · outbound

This paper cites Rethinking the learning paradigm for dynamic facial expression recognition.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Rethinking the learning paradigm for dynamic facial expression recognition

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.831395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.218238Z digest=sha256:44d6a83d3ab138548de5c1c59d46d4554f0f94b86814147f7ef100ad11e22737

Observation 35934a95-d9e1-42a0-a6dc-1c0ce49d7288 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.223723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.223723Z digest=sha256:cf5f9371c57da70434cea676351f28269eafc687d2b89287d8cad19a847d6338

Observation de36e934-6782-42fa-9aee-d879762dbda5 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.233811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.233811Z digest=sha256:8acd23dcfb85cd54b93a94c315d79bc3dd1dc6e21491a4180495c9dd574f3950

Observation f76c93ae-0f6e-42bc-ae1f-c6ae17b41707 · outbound

This paper cites Emovit: Revolutionizing emotion insights with visual instruction tuning.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Emovit: Revolutionizing emotion insights with visual instruction tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.240611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.240611Z digest=sha256:caae065ad4610842ff43e8f0565a7bd22bb282e315629b5e50bfac1a05649406

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.247106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.247106Z digest=sha256:629d38ba4497152790692d55f9fc425a0cf806ede30eee72c547c92fa77362c0

Observation c69de7ab-bd21-4e78-969f-d0d6795b13a4 · outbound

This paper cites Robust emotion recognition in context debiasing.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Robust emotion recognition in context debiasing

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.793904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.255486Z digest=sha256:687160096806f51978afbab46e2c122c084d7cba32cc69355d6393fcabe92f31

Observation 492de562-dca7-4574-bf19-7a55fa8adf0d · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.262550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.262550Z digest=sha256:2dad19e85ead8dd745d4a5557abe03256c06f9b57d2fff7af922e16a07ef36b0

Observation aebe6239-57eb-4344-ada0-bd843454f20a · outbound

This paper cites Sigmoid loss for language image pre-training.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Sigmoid loss for language image pre-training

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.276459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.276459Z digest=sha256:3441f25f74a5dcb7b505fd012c5d34eb308bfc031d89aaa2bea7a106c0b3b216

Observation c66b5d0a-dbb2-4932-9697-f7fb8ebd4415 · outbound

This paper cites A high-resolution spontaneous 3d dynamic facial expression database.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation A high-resolution spontaneous 3d dynamic facial expression database

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.761442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.284777Z digest=sha256:051c074cf1050b979e3d09c0609ed1ac0ce32c94ac03039fc902cad0e70f3f68

Observation 45d9dd06-f5d7-46be-809d-79e895bd6a15 · outbound

This paper cites Bp4d-spontaneous: a high-resolution spontaneous 3d dynamic facial expression database.Image and Vision Computing, 32(10):692–706, 2014.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Bp4d-spontaneous: a high-resolution spontaneous 3d dynamic facial expression database.Image and Vision Computing, 32(10):692–706, 2014

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:44:29.290214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:44:29.290214Z digest=sha256:319336b45845ffa94468c4d99a1c4410abd5134fb757780739ed6ac495c92b0e

Observation d99c2338-975d-412d-81a8-97d9fad1dd0c · outbound

This paper cites Khfa: Knowledge-driven hierarchical feature alignment framework for subject- invariant facial action unit detection.IEEE Transactions on Instrumentation and Measurement, 2024.

Emotion Knowledge Enhancement for Vision Large Language Models: A Self-Verification Approach for High-Quality Emotion Instruction Data Generation Khfa: Knowledge-driven hierarchical feature alignment framework for subject- invariant facial action unit detection.IEEE Transactions on Instrumentation and Measurement, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:44:29.730776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:44:29.296392Z digest=sha256:c4a7943b2e28565d501fbaf12e6fb67fbcde17a5800120d1daa3279da5964fce

Pith citing papers

No inbound Pith citation observations are available.