Pith. sign in

Paper Citation Record · LEDGER

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues

As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2502.00397.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00397 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:13:03.461605Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy38
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e431000-728e-4e1b-aea8-b111e21bf261 · outbound

This paper cites Visual saliency model for robot cameras,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Visual saliency model for robot cameras,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.847907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.340855Z digest=sha256:bfd671ad11f08cf9d34283ab227f35877dd204f24f559a797e63218fd9907dc8

Observation ea8f4b89-fb3f-48f3-872b-0237b650bcf2 · outbound

This paper cites Gazed– gaze-guided cinematic editing of wide-angle monocular video record- ings,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Gazed– gaze-guided cinematic editing of wide-angle monocular video record- ings,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.839453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.344735Z digest=sha256:79926bc6005413c1e27900dc9dfe481dc55e48703cac5cc82886bff872c885df

Observation ebdca697-77f8-48eb-ac5b-2f3c87437f02 · outbound

This paper cites Salgaze: Personalizing gaze estimation using visual saliency,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Salgaze: Personalizing gaze estimation using visual saliency,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.830304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.348390Z digest=sha256:67ce74d3193c803eb13d3b0e4e08304f8f78a643c9b4aaf9dfb3b02d9d69e34a

Observation c10f5d01-7c3b-446e-b2af-0623aca4cfa0 · outbound

This paper cites Attentional mechanisms for socially in- teractive robots–a survey,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Attentional mechanisms for socially in- teractive robots–a survey,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.821962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.352217Z digest=sha256:dcad3940f4be81a88fee99ac15cfec0fc7fc635d4f28010518a68a2cec48091b

Observation af87c647-ecca-491f-ae5b-d0842646a5cf · outbound

This paper cites Facial expression recognition using visual saliency and deep learning,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Facial expression recognition using visual saliency and deep learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.811740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.355918Z digest=sha256:ab55f63692146f65baceaaec792160f36c0dc83a772dc6d952221f82dd26133a

Observation 49e7d367-ff0e-4b54-b843-8fda5831d9a7 · outbound

This paper cites Evaluating the effect of saliency detection and attention manipulation in human-robot interac- tion,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Evaluating the effect of saliency detection and attention manipulation in human-robot interac- tion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.800927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.358976Z digest=sha256:3f29b0787f57e4da6c8879cd0e9788a6ad892679d29dbdd199545f44bb3398c6

Observation 71c974f7-5efd-4a6d-85ad-69b6863326c5 · outbound

This paper cites Saliency heat-map as visual attention for autonomous driving using generative adversarial network (gan),.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Saliency heat-map as visual attention for autonomous driving using generative adversarial network (gan),

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.792841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.362398Z digest=sha256:b294328a07ba48adfff82be71b0a31e27647d2775c39eb0d338436758505dfb4

Observation 1177dafd-542c-4c6f-8253-2f650bdd1c93 · outbound

This paper cites A gated fusion network for dynamic saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues A gated fusion network for dynamic saliency prediction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.784138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.365227Z digest=sha256:929b568cf0df36b1b26e94178ed18418de35335f76820cb5e9dfc35b8eadc1c0

Observation 05b6f720-53b8-4248-8eb6-5b3fb12eaede · outbound

This paper cites Video saliency prediction based on spatial- temporal two-stream network,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Video saliency prediction based on spatial- temporal two-stream network,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.775208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.368512Z digest=sha256:5354b7a12d7ff7d9954a95d3ba189f311dc3e4105ea1af5fd2a5eb619e621fe8

Observation ffb2f9a9-ce96-4b5a-859a-a959d8ef6a78 · outbound

This paper cites Unified image and video saliency modeling,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Unified image and video saliency modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.765839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.371282Z digest=sha256:8497466287f5f29cd729debc03759a7c3a6d234341d133f6cb38147cba5bec3e

Observation 9f1fac0d-049c-4212-a070-9d84695b0d97 · outbound

This paper cites Revisiting video saliency prediction in the deep learning era,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Revisiting video saliency prediction in the deep learning era,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.754282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.374266Z digest=sha256:0207a0ad16ac111dc3e77e6120a1d2b26928e7e977ff82f0041785e06eeb90b1

Observation a8c4f7a6-d2e7-4430-b913-58bfb2b37ddb · outbound

This paper cites Vinet: Pushing the limits of visual modality for audio-visual saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Vinet: Pushing the limits of visual modality for audio-visual saliency prediction,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.744643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.377452Z digest=sha256:2d660ff12b7ce8cc8f08303a352d14b6808d5fd556f4ec1f0e4343f4f947cbf8

Observation a605d4ca-b65b-4870-8f45-5093569f68a6 · outbound

This paper cites Tased-net: Temporally-aggregating spatial encoder-decoder network for video saliency detection,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Tased-net: Temporally-aggregating spatial encoder-decoder network for video saliency detection,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.735251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.381064Z digest=sha256:29acacb125789830ab074a1840fcd4c6d3de5ffa4fc8b445641072aabdf6642e

Observation bcb8e4b9-a45e-4e35-ad53-6b05290b8ef3 · outbound

This paper cites Rethinking spatiotem- poral feature learning: Speed-accuracy trade-offs in video classification,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Rethinking spatiotem- poral feature learning: Speed-accuracy trade-offs in video classification,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.725740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.383843Z digest=sha256:14a8fecb4ce8161f4478b56e8fd241eebe2fc2c4e5438c35025d04c6692604ff

Observation 419474bc-90d4-42ff-9124-9fdff43a794f · outbound

This paper cites The Kinetics Human Action Video Dataset.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues The Kinetics Human Action Video Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T19:13:03.386650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:13:03.386650Z digest=sha256:9aec8df8e19f67319ede4c22e3f398453b627b5a7c9d3e65377155e824fc3baf

Observation 872b3164-3bfb-4a17-8472-3d7ef398fe1f · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues U-net: Convolutional networks for biomedical image segmentation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T19:13:03.390295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:13:03.390295Z digest=sha256:d83956476e3dcbaa0058720362ee87617af07fad24780754c16eca3ed39dd276

Observation b76305d9-f5ee-4e99-8d03-eedbb9d4ce74 · outbound

This paper cites Spatio-temporal self-attention network for video saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Spatio-temporal self-attention network for video saliency prediction,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.710799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.393113Z digest=sha256:586285d8077e71e57b138518dbc57c25f7274c37847b9eb8b4705ad4a04b85fe

Observation 51ff5cc8-b091-49f0-b5e6-b401ac11c645 · outbound

This paper cites Transformer-based multi-scale feature integration network for video saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Transformer-based multi-scale feature integration network for video saliency prediction,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.701581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.395737Z digest=sha256:15b3786b2285af494af58b111456e43ef3659cda35eb0afeee9c3aab055f8b3e

Observation 911fb5e8-7de5-4820-81b0-0fefad5798d4 · outbound

This paper cites Transformer-based video saliency prediction with high temporal dimension decoding,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Transformer-based video saliency prediction with high temporal dimension decoding,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.692733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.398636Z digest=sha256:39f052379edd4fe6adb0fe2616abe033e0a83a6a0c3feac100f6ee22e020c70b

Observation afa4fb62-e2a1-44a5-a7dd-74963c6c938c · outbound

This paper cites Stavis: Spatio-temporal audio- visual saliency network,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Stavis: Spatio-temporal audio- visual saliency network,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.682802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.401401Z digest=sha256:5401b87056832a55f333ce535cc65b87d3439a083848da2dce8a85dd21fb1b55

Observation 1fb34a7e-f573-4d0a-b98e-6a4f6650a71b · outbound

This paper cites Temporal-Spatial Feature Pyramid for Video Saliency Detection.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Temporal-Spatial Feature Pyramid for Video Saliency Detection

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T19:13:03.404771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:13:03.404771Z digest=sha256:04ca2f93901aa615e6cbaace8421eeee16e4c6949956f5c1a49490c6f985d0fb

Observation 70832a52-b68e-4438-a255-ab88e0ae3b44 · outbound

This paper cites Joint learning of audio-visual saliency prediction and sound source localization on multi-face videos,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Joint learning of audio-visual saliency prediction and sound source localization on multi-face videos,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.673464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.407912Z digest=sha256:9680de42d9273b881366b4239b5224c752524aeecb87889f9490d8bedb9e6882

Observation 10e67ec7-b294-463b-9d07-8daf50c0031e · outbound

This paper cites Learning to predict salient faces: A novel visual-audio saliency model,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Learning to predict salient faces: A novel visual-audio saliency model,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.665172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.410741Z digest=sha256:30128b2abd0376141ab9cb60b38275ec075d4874944e06afb188980840ea0a17

Observation 6b3a2fc4-d9c1-4727-87c3-37a87d911c38 · outbound

This paper cites Casp-net: Rethinking video saliency prediction from an audio-visual consistency perceptual perspective,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Casp-net: Rethinking video saliency prediction from an audio-visual consistency perceptual perspective,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.657150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.413446Z digest=sha256:0761c92e0445e2a61146341b2e16023f685a486c67ca9af3866c0a1fc69399c3

Observation b1889d04-22b4-49b3-a7e1-9227b9f3123d · outbound

This paper cites Diffsal: Joint audio and video learning for diffusion saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Diffsal: Joint audio and video learning for diffusion saliency prediction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.648146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.416329Z digest=sha256:9dbb405fd1a974d9eda602dc8ecb8a5f4e46011eb53d5171d23044d0a20c501e

Observation a4b3a613-a75d-409a-9d61-a63bc0025d99 · outbound

This paper cites Deep roots: Improving cnn efficiency with hierarchical filter groups,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Deep roots: Improving cnn efficiency with hierarchical filter groups,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.639371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.419165Z digest=sha256:b83a1c17aa1bd6dde3c95bbb0132d3f2e6b1a906141a3c6c45f714eb21b32fd0

Observation 38379fce-13a5-4d5c-bc89-878ef2f644fb · outbound

This paper cites Shufflenet: An extremely efficient convolutional neural network for mobile devices,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Shufflenet: An extremely efficient convolutional neural network for mobile devices,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.631380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.421867Z digest=sha256:8d7b1aacdfe32ec201346c5bfaaa9ed8173e4532ad9fe6890550e4f4764d82e5

Observation e9d9ee8c-c587-4518-b985-b5ab1e8436dd · outbound

This paper cites Actor- context-actor relation network for spatio-temporal action localization,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Actor- context-actor relation network for spatio-temporal action localization,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.623245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.424569Z digest=sha256:1754b1736d2729ff3f0f6cdccc7d5b69f60bcd6c1349ce491415aeef59274bd6

Observation 4e26c035-8412-4faa-9f0b-278113f4b862 · outbound

This paper cites Slowfast networks for video recognition,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Slowfast networks for video recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.613895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.427677Z digest=sha256:ddd5c572c10b80b1c4d411b652bc20d13ab4e1ee64efdc6df194310a6697188d

Observation fd2f8970-3785-4b1f-80c7-85c8171a254e · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Ava: A video dataset of spatio-temporally localized atomic visual actions,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.605955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.430271Z digest=sha256:3215a90d33e17f62740c7d252f07f812e628dccd8c8fbc32d8fad5950b9fc839

Observation 8cb2d9bb-038a-4f02-b8f1-5b6f8f2af615 · outbound

This paper cites Actions in the eye: Dynamic gaze datasets and learnt saliency models for visual recognition,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Actions in the eye: Dynamic gaze datasets and learnt saliency models for visual recognition,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.597676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.432862Z digest=sha256:e5accad0dc991dd400d714fd1b1c7cc8e6ee4bfea9bd715cfe3108e34be60e85

Observation 52e2a6d1-abd0-4b70-8af4-0adb6a543475 · outbound

This paper cites Fixation prediction through multimodal analysis,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Fixation prediction through multimodal analysis,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.589435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.435844Z digest=sha256:5ac8281d8f91b754a2c2c0e29ca3e16be112eb6f4c8a91f0e24349e617fe12e6

Observation 629d8538-fa1b-4492-acef-c46d43f675eb · outbound

This paper cites How saliency, faces, and sound influence gaze in dynamic social scenes,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues How saliency, faces, and sound influence gaze in dynamic social scenes,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.579363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.439170Z digest=sha256:8af1f224cb2306e77c51ed7d9405ccf4834f9686718971b6412033379e4e4265

Observation 0c4e7ff8-ff2c-4f33-9122-41949582ce8c · outbound

This paper cites Toward the introduction of auditory information in dynamic visual attention models,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Toward the introduction of auditory information in dynamic visual attention models,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.569696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.441720Z digest=sha256:6a1e0f210b95d2c750959f76320c89ca067dab158ecaf02e24f6338f85c79a64

Observation 3776c107-34d1-456c-9e00-fb98a76caced · outbound

This paper cites An efficient audiovisual saliency model to predict eye positions when looking at conversations,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues An efficient audiovisual saliency model to predict eye positions when looking at conversations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.559930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.444399Z digest=sha256:def6f9eefd213ca11e31edd60aef775cf83cb7083dc62fe5b1c4022897b372ee

Observation f3f96374-e6b4-4a69-b076-9742178c10c0 · outbound

This paper cites Clustering of gaze during dynamic scene viewing is predicted by motion,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Clustering of gaze during dynamic scene viewing is predicted by motion,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.549618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.447075Z digest=sha256:034adfa9fe6108c19b2f2e678322353a41a2a15ab2f3455c1f87f4462c131f60

Observation 3fe200b4-4cd8-4d74-8c52-66005e90d132 · outbound

This paper cites Predicting eyes’ fixa- tions in movie videos: Visual saliency experiments on a new eye- tracking database,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Predicting eyes’ fixa- tions in movie videos: Visual saliency experiments on a new eye- tracking database,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.539925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.450023Z digest=sha256:23278d8d7952401626687e2cdcb5c37addfaf37b60b68e07ace432f8b331be37

Observation 4a8c7cbd-e557-4448-8782-5197ce51d1fd · outbound

This paper cites Tinyhd: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Tinyhd: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.530709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.452673Z digest=sha256:7423fbbe890aaa737aaadf35e5c4d2abecbdffd2b8117e54994fcaddaf111c6d

Observation c0e331bd-fbe9-4858-a4ff-4d6849cba445 · outbound

This paper cites Video saliency forecasting transformer,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Video saliency forecasting transformer,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.520968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.455473Z digest=sha256:d1104faaf570dd197246319cb52ce6d571d521377c00a494971e8e194d73f265

Observation 6a52a3c4-8d4b-4839-b001-4c5fd5744ede · outbound

This paper cites What do different evaluation metrics tell us about saliency models?.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues What do different evaluation metrics tell us about saliency models?

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.512313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.458030Z digest=sha256:a60c5320220ef05bda80b3b63661eeb05c582f30d9a41aaa36a4bd0feddb970a

Observation 0dd95823-6a1f-4c1c-925a-c30a2d44240b · outbound

This paper cites Does audio help in deep audio-visual saliency prediction models?.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Does audio help in deep audio-visual saliency prediction models?

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.501771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-09T19:13:03.461605Z digest=sha256:2d56da972208785f1144f8a917fcc5c967af257b4a2a66027766a325956cb558

Pith citing papers

No inbound Pith citation observations are available.