Pith. sign in

Paper Citation Record · LEDGER

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition

As of 10 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2607.09417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.09417 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-13T03:10:38.745341Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 667f7e2c-ed4f-4fb7-aaad-5ed1598c46b4 · outbound

This paper cites Multimodal Ambivalence/Hesitancy Recognition in Videos for Personalized Digital Health Interventions.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Multimodal Ambivalence/Hesitancy Recognition in Videos for Personalized Digital Health Interventions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:43756b9559603391958df623d943861ae14a998812c8d3f0db95a49acfbfd7a7

Observation 8d43a031-0a57-4ec4-991a-1e1ca42b84a6 · outbound

This paper cites Nonverbal behavior in clinician—patient interaction.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Nonverbal behavior in clinician—patient interaction

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:f0a43195754e5d813caa1bc312ed450aec7abd0c00e5dcd9f855cb3fe28e237c

Observation 049c35d1-401e-43e0-bc98-1d6c7f14ae20 · outbound

This paper cites Methods to assess ambivalence towards food and diet: a scoping review.Journal of Human Nutrition and Dietetics, 36(5): 2010–2025, 2023.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Methods to assess ambivalence towards food and diet: a scoping review.Journal of Human Nutrition and Dietetics, 36(5): 2010–2025, 2023

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:43c3fdf48abd7c1160fa5f287077c4c797764870767b4a2ca978b8286dd41a4f

Observation f39461fc-12d2-43ec-8664-fada5fcc43e9 · outbound

This paper cites Understanding and predicting health behaviour change: a contemporary view through the lenses of meta-reviews.Health psychology review, 14(1):1–5, 2020.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Understanding and predicting health behaviour change: a contemporary view through the lenses of meta-reviews.Health psychology review, 14(1):1–5, 2020

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:fb180e8596e96b2ddf35f4aac8ec2c75d63a58a0a58123c98cce32866680f311

Observation e7878cd9-2bec-447e-9612-f28fe3c792af · outbound

This paper cites Multi-label compound expression recognition: C-expr database & network.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Multi-label compound expression recognition: C-expr database & network

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:0b84e5693a1420e9501d87f4b02f4a2ed4e3b97b43ef2f1248a64f119e6013da

Observation 30536d54-2016-41ee-88a6-852268e18e5b · outbound

This paper cites BAH Dataset for Ambivalence/Hesitancy Recognition in Videos for Digital Behavioural Change.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition BAH Dataset for Ambivalence/Hesitancy Recognition in Videos for Digital Behavioural Change

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:2dd1845dd43d21c7f8d3e897635b196b3e634cc8bad17ca7c3358b2cfd7b32bf

Observation a80de8b0-ab01-4ccb-84c4-ee8f1fe08715 · outbound

This paper cites How children and adults produce and perceive uncertainty in audiovisual speech.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition How children and adults produce and perceive uncertainty in audiovisual speech

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:e83d2e15232ef178fd8c5b30efe2f9beae671a888555b7780eb45c8a899622c5

Observation 5b7a7558-75b7-47bd-a4d1-2af60a276ef3 · outbound

This paper cites Toward an affect-sensitive multimodal human-computer interaction.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Toward an affect-sensitive multimodal human-computer interaction

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:a19daa09bf0046091ece511de9181a866866ff029087eb044da518d70a879f7c

Observation 85984eca-9327-403e-8f10-fe87018a19ca · outbound

This paper cites Emotion recognition from multiple modalities: Fundamentals and methodologies.IEEE Signal Processing Magazine, 38(6):59–73, 2021.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Emotion recognition from multiple modalities: Fundamentals and methodologies.IEEE Signal Processing Magazine, 38(6):59–73, 2021

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:d33cbe898ad1e4f3ed81a0f8215ef8c2b8d38e4b2608dbc1dd2b2e663ed22691

Observation 25715ef6-4d53-430a-8020-a58fab1c8cbf · outbound

This paper cites Aligning multimodal data for fine-grained video understanding via cross- attentive recurrent fusion.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Aligning multimodal data for fine-grained video understanding via cross- attentive recurrent fusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:f08a1d8d07978d7a026cafd60a2860be249de2999528ed6391c74565843cf3dd

Observation b71a8aed-ff28-429e-93ef-8b557d4163d8 · outbound

This paper cites Multimodal spontaneous emotion corpus for human behavior analysis.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Multimodal spontaneous emotion corpus for human behavior analysis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:9a88fa3a07cc990ca7b9109507756a8ac675083223199da7e72496979a8726db

Observation 6381f075-6015-4bfe-b150-e1beba57f4d2 · outbound

This paper cites Are multimodal transformers robust to missing modality? InProceedings of the IEEE/CVF conference on computer vision and pattern recognition, pages 18177–18186, 2022.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Are multimodal transformers robust to missing modality? InProceedings of the IEEE/CVF conference on computer vision and pattern recognition, pages 18177–18186, 2022

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:0b244194c5fcaed72330e35576a9a4f17f953eeb0d114aafae416aa952434ee8

Observation 3f9aa0e4-4e7c-4594-8659-657749348971 · outbound

This paper cites Hidden emotion detection using multi-modal signals.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Hidden emotion detection using multi-modal signals

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:2dc1b7704d1cdd8de9cae23f61f2c06986e0d12d46253f101de65ac3bed5ec1c

Observation 59aacab1-54de-401a-8203-c51ce914126c · outbound

This paper cites Triagedmsa: Triaging sentimental disagreement in multimodal sentiment analysis.IEEE transactions on affective computing, 16(3): 1557–1569, 2025.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Triagedmsa: Triaging sentimental disagreement in multimodal sentiment analysis.IEEE transactions on affective computing, 16(3): 1557–1569, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:52f2f5dcc3cb60d656b8abe4727b04647c6ed704afcd357f0519410c3372e829

Observation fb0802ec-e603-4c2d-9b34-0713d5b93495 · outbound

This paper cites General facial representation learning in a visual-linguistic manner.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition General facial representation learning in a visual-linguistic manner

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:c6a3a733f2d5b25c0ee3580eabb5ae8234f3bd57432533f76e6f97390385f7d0

Observation 60ea1ad6-2772-49da-8cea-ede39cfabe23 · outbound

This paper cites Crossvit: Cross-attention multi-scale vision trans- former for image classification.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Crossvit: Cross-attention multi-scale vision trans- former for image classification

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:88c86b7249078f3f9992367e96b46bfee9581882a089c664d4464826aa84d07a

Observation 18fe05f9-cb0e-4c6c-88e6-d19bf99ada4f · outbound

This paper cites Affective behavior analysis in-the-wild challenge, 2026.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Affective behavior analysis in-the-wild challenge, 2026

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:f42f8a0599e11c0b501c736aa19586e7914c496811c88bdf9f597fb0ac04543b

Observation d467f90a-7a15-4aa6-8063-4bf9782cc323 · outbound

This paper cites Enhanced lstm for natural language inference.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Enhanced lstm for natural language inference

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:a3a01d940ef8a491d6b2c5019d306a9e04ecdcc6c0824bacd7eeeace508bac95

Observation 8c3272ea-10e5-43d5-95f9-503efacaed69 · outbound

This paper cites Qwen technical report, 2023.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Qwen technical report, 2023

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:b7f7ea9507a7b10af991a7df0cdaeea85fde5c2bc18e1c122211aabc201009bd

Observation 8e2ea022-1d43-4ef2-b26d-7fadfa0f8178 · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.Advances in Neural Information Processing Systems, 2022.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.Advances in Neural Information Processing Systems, 2022

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:78fad3ce2b2c03d1df24d6ed78159905febb8f68084000a9e020dd09e4bbb2e9

Observation dad2d7f1-33ee-4a44-bdd7-395f3ff6db84 · outbound

This paper cites Robust speech recognition via large-scale weak supervision.

SVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition Robust speech recognition via large-scale weak supervision

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-13T03:10:38.745341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:10:38.745341Z digest=sha256:ec30819e39b3c702a447550a902838e8dd94155dbd1fd1d3c592e1911bb4560b

Pith citing papers

No inbound Pith citation observations are available.