Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T19:13:03.461605Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2502.00397.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T19:13:03.461605Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6e431000-728e-4e1b-aea8-b111e21bf261 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Visual saliency model for robot cameras,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ea8f4b89-fb3f-48f3-872b-0237b650bcf2 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Gazed– gaze-guided cinematic editing of wide-angle monocular video record- ings,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ebdca697-77f8-48eb-ac5b-2f3c87437f02 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Salgaze: Personalizing gaze estimation using visual saliency,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c10f5d01-7c3b-446e-b2af-0623aca4cfa0 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Attentional mechanisms for socially in- teractive robots–a survey,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation af87c647-ecca-491f-ae5b-d0842646a5cf · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Facial expression recognition using visual saliency and deep learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 49e7d367-ff0e-4b54-b843-8fda5831d9a7 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Evaluating the effect of saliency detection and attention manipulation in human-robot interac- tion,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 71c974f7-5efd-4a6d-85ad-69b6863326c5 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Saliency heat-map as visual attention for autonomous driving using generative adversarial network (gan),
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1177dafd-542c-4c6f-8253-2f650bdd1c93 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues A gated fusion network for dynamic saliency prediction,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 05b6f720-53b8-4248-8eb6-5b3fb12eaede · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Video saliency prediction based on spatial- temporal two-stream network,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ffb2f9a9-ce96-4b5a-859a-a959d8ef6a78 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Unified image and video saliency modeling,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9f1fac0d-049c-4212-a070-9d84695b0d97 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Revisiting video saliency prediction in the deep learning era,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a8c4f7a6-d2e7-4430-b913-58bfb2b37ddb · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Vinet: Pushing the limits of visual modality for audio-visual saliency prediction,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a605d4ca-b65b-4870-8f45-5093569f68a6 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Tased-net: Temporally-aggregating spatial encoder-decoder network for video saliency detection,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bcb8e4b9-a45e-4e35-ad53-6b05290b8ef3 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Rethinking spatiotem- poral feature learning: Speed-accuracy trade-offs in video classification,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 419474bc-90d4-42ff-9124-9fdff43a794f · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues The Kinetics Human Action Video Dataset
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 872b3164-3bfb-4a17-8472-3d7ef398fe1f · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues U-net: Convolutional networks for biomedical image segmentation,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b76305d9-f5ee-4e99-8d03-eedbb9d4ce74 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Spatio-temporal self-attention network for video saliency prediction,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 51ff5cc8-b091-49f0-b5e6-b401ac11c645 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Transformer-based multi-scale feature integration network for video saliency prediction,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 911fb5e8-7de5-4820-81b0-0fefad5798d4 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Transformer-based video saliency prediction with high temporal dimension decoding,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation afa4fb62-e2a1-44a5-a7dd-74963c6c938c · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Stavis: Spatio-temporal audio- visual saliency network,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1fb34a7e-f573-4d0a-b98e-6a4f6650a71b · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Temporal-Spatial Feature Pyramid for Video Saliency Detection
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70832a52-b68e-4438-a255-ab88e0ae3b44 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Joint learning of audio-visual saliency prediction and sound source localization on multi-face videos,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 10e67ec7-b294-463b-9d07-8daf50c0031e · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Learning to predict salient faces: A novel visual-audio saliency model,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6b3a2fc4-d9c1-4727-87c3-37a87d911c38 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Casp-net: Rethinking video saliency prediction from an audio-visual consistency perceptual perspective,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b1889d04-22b4-49b3-a7e1-9227b9f3123d · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Diffsal: Joint audio and video learning for diffusion saliency prediction,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4b3a613-a75d-409a-9d61-a63bc0025d99 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Deep roots: Improving cnn efficiency with hierarchical filter groups,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 38379fce-13a5-4d5c-bc89-878ef2f644fb · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Shufflenet: An extremely efficient convolutional neural network for mobile devices,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e9d9ee8c-c587-4518-b985-b5ab1e8436dd · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Actor- context-actor relation network for spatio-temporal action localization,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4e26c035-8412-4faa-9f0b-278113f4b862 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Slowfast networks for video recognition,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fd2f8970-3785-4b1f-80c7-85c8171a254e · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Ava: A video dataset of spatio-temporally localized atomic visual actions,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8cb2d9bb-038a-4f02-b8f1-5b6f8f2af615 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Actions in the eye: Dynamic gaze datasets and learnt saliency models for visual recognition,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 52e2a6d1-abd0-4b70-8af4-0adb6a543475 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Fixation prediction through multimodal analysis,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 629d8538-fa1b-4492-acef-c46d43f675eb · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues How saliency, faces, and sound influence gaze in dynamic social scenes,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0c4e7ff8-ff2c-4f33-9122-41949582ce8c · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Toward the introduction of auditory information in dynamic visual attention models,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3776c107-34d1-456c-9e00-fb98a76caced · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues An efficient audiovisual saliency model to predict eye positions when looking at conversations,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f3f96374-e6b4-4a69-b076-9742178c10c0 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Clustering of gaze during dynamic scene viewing is predicted by motion,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3fe200b4-4cd8-4d74-8c52-66005e90d132 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Predicting eyes’ fixa- tions in movie videos: Visual saliency experiments on a new eye- tracking database,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4a8c7cbd-e557-4448-8782-5197ce51d1fd · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Tinyhd: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c0e331bd-fbe9-4858-a4ff-4d6849cba445 · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Video saliency forecasting transformer,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6a52a3c4-8d4b-4839-b001-4c5fd5744ede · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues What do different evaluation metrics tell us about saliency models?
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0dd95823-6a1f-4c1c-925a-c30a2d44240b · outbound
Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Does audio help in deep audio-visual saliency prediction models?
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.