Pith. sign in

Paper Citation Record · LEDGER

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues

As of 11 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2502.00397.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00397 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:13:03.461605Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy38
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6e431000-728e-4e1b-aea8-b111e21bf261 · outbound

This paper cites Visual saliency model for robot cameras,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Visual saliency model for robot cameras,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.847907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.340855Z digest=sha256:1f309826d61b0b461646f4a191ba9254656f8e13ae7cead722591d8a8ed47c65

Observation ea8f4b89-fb3f-48f3-872b-0237b650bcf2 · outbound

This paper cites Gazed– gaze-guided cinematic editing of wide-angle monocular video record- ings,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Gazed– gaze-guided cinematic editing of wide-angle monocular video record- ings,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.839453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.344735Z digest=sha256:40828e5869d2601083beb0f19d7b06a060f839e2b9ca968aa5b63344ab864039

Observation ebdca697-77f8-48eb-ac5b-2f3c87437f02 · outbound

This paper cites Salgaze: Personalizing gaze estimation using visual saliency,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Salgaze: Personalizing gaze estimation using visual saliency,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.830304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.348390Z digest=sha256:b5e8f25827451b58ee06daa33b4ccf6b91d96899b87e133cfb74f7486db7c8a7

Observation c10f5d01-7c3b-446e-b2af-0623aca4cfa0 · outbound

This paper cites Attentional mechanisms for socially in- teractive robots–a survey,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Attentional mechanisms for socially in- teractive robots–a survey,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.821962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.352217Z digest=sha256:1716910a155e2d96fd610bd65d4a3f99a39b8ebd4d03ecff5aeddfd79993855f

Observation af87c647-ecca-491f-ae5b-d0842646a5cf · outbound

This paper cites Facial expression recognition using visual saliency and deep learning,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Facial expression recognition using visual saliency and deep learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.811740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.355918Z digest=sha256:c94a3781f702c7b88c4a81b47b317d229ccb56c627463c8a7642147a648f6c00

Observation 49e7d367-ff0e-4b54-b843-8fda5831d9a7 · outbound

This paper cites Evaluating the effect of saliency detection and attention manipulation in human-robot interac- tion,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Evaluating the effect of saliency detection and attention manipulation in human-robot interac- tion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.800927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.358976Z digest=sha256:c943c20c3e65c840e17fbecd15dd41049b577d6300363ae88ff4ffa102258207

Observation 71c974f7-5efd-4a6d-85ad-69b6863326c5 · outbound

This paper cites Saliency heat-map as visual attention for autonomous driving using generative adversarial network (gan),.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Saliency heat-map as visual attention for autonomous driving using generative adversarial network (gan),

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.792841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.362398Z digest=sha256:0bff64f6bddc9981ef53f966885408d35c45ca30362ab729d55121ec2454c75a

Observation 1177dafd-542c-4c6f-8253-2f650bdd1c93 · outbound

This paper cites A gated fusion network for dynamic saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues A gated fusion network for dynamic saliency prediction,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.784138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.365227Z digest=sha256:ae781c328f8f1a7fc565eb7e1a5b3612630718661caca13f4e4c95596b5653ad

Observation 05b6f720-53b8-4248-8eb6-5b3fb12eaede · outbound

This paper cites Video saliency prediction based on spatial- temporal two-stream network,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Video saliency prediction based on spatial- temporal two-stream network,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.775208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.368512Z digest=sha256:d1ffc1d938b73909d821cd73903849fc09094af7a069624e5b5eb9717275560d

Observation ffb2f9a9-ce96-4b5a-859a-a959d8ef6a78 · outbound

This paper cites Unified image and video saliency modeling,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Unified image and video saliency modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.765839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.371282Z digest=sha256:83667f3638d04185c3721a71158b3b5dd09a187fa10e79a69a63e377b86c78e7

Observation 9f1fac0d-049c-4212-a070-9d84695b0d97 · outbound

This paper cites Revisiting video saliency prediction in the deep learning era,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Revisiting video saliency prediction in the deep learning era,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.754282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.374266Z digest=sha256:171883694fb369e82648066f5d0559c231c2e7d128a39d6eb626c9469f1584fa

Observation a8c4f7a6-d2e7-4430-b913-58bfb2b37ddb · outbound

This paper cites Vinet: Pushing the limits of visual modality for audio-visual saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Vinet: Pushing the limits of visual modality for audio-visual saliency prediction,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.744643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.377452Z digest=sha256:09a7824905af247025516450ae6b7c392044bad3b7e1f374edf04f38c4b0ba1f

Observation a605d4ca-b65b-4870-8f45-5093569f68a6 · outbound

This paper cites Tased-net: Temporally-aggregating spatial encoder-decoder network for video saliency detection,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Tased-net: Temporally-aggregating spatial encoder-decoder network for video saliency detection,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.735251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.381064Z digest=sha256:5f7ca3b6b461dcb6cd96310a5ae22947519210fba917ab367e78567b241962d9

Observation bcb8e4b9-a45e-4e35-ad53-6b05290b8ef3 · outbound

This paper cites Rethinking spatiotem- poral feature learning: Speed-accuracy trade-offs in video classification,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Rethinking spatiotem- poral feature learning: Speed-accuracy trade-offs in video classification,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.725740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.383843Z digest=sha256:706f13949d4ba248bf5a18b5e3930d0dbcfc3e6abf0ed545dc74738f35938a9f

Observation 419474bc-90d4-42ff-9124-9fdff43a794f · outbound

This paper cites The Kinetics Human Action Video Dataset.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues The Kinetics Human Action Video Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T19:13:03.386650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:13:03.386650Z digest=sha256:9aec8df8e19f67319ede4c22e3f398453b627b5a7c9d3e65377155e824fc3baf

Observation 872b3164-3bfb-4a17-8472-3d7ef398fe1f · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues U-net: Convolutional networks for biomedical image segmentation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T19:13:03.390295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:13:03.390295Z digest=sha256:d83956476e3dcbaa0058720362ee87617af07fad24780754c16eca3ed39dd276

Observation b76305d9-f5ee-4e99-8d03-eedbb9d4ce74 · outbound

This paper cites Spatio-temporal self-attention network for video saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Spatio-temporal self-attention network for video saliency prediction,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.710799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.393113Z digest=sha256:938f38ecfa933928b63fbcf86f4b7da3e8ee85271a06a7297e036ce4916f0585

Observation 51ff5cc8-b091-49f0-b5e6-b401ac11c645 · outbound

This paper cites Transformer-based multi-scale feature integration network for video saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Transformer-based multi-scale feature integration network for video saliency prediction,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.701581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.395737Z digest=sha256:9be18e0cd0be768847648c7d4993c589a74181544b87e4168ef74d1cfc968549

Observation 911fb5e8-7de5-4820-81b0-0fefad5798d4 · outbound

This paper cites Transformer-based video saliency prediction with high temporal dimension decoding,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Transformer-based video saliency prediction with high temporal dimension decoding,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.692733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.398636Z digest=sha256:fe4960c8012fa4d9c1c5bfff31ebf0904f364223e95476f929bdb1ffc285585b

Observation afa4fb62-e2a1-44a5-a7dd-74963c6c938c · outbound

This paper cites Stavis: Spatio-temporal audio- visual saliency network,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Stavis: Spatio-temporal audio- visual saliency network,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.682802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.401401Z digest=sha256:d796669bdc3a9abb68973c375675aab07129f4bf8e13a91a8530c1bbcf3885b7

Observation 1fb34a7e-f573-4d0a-b98e-6a4f6650a71b · outbound

This paper cites Temporal-Spatial Feature Pyramid for Video Saliency Detection.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Temporal-Spatial Feature Pyramid for Video Saliency Detection

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T19:13:03.404771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:13:03.404771Z digest=sha256:04ca2f93901aa615e6cbaace8421eeee16e4c6949956f5c1a49490c6f985d0fb

Observation 70832a52-b68e-4438-a255-ab88e0ae3b44 · outbound

This paper cites Joint learning of audio-visual saliency prediction and sound source localization on multi-face videos,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Joint learning of audio-visual saliency prediction and sound source localization on multi-face videos,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.673464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.407912Z digest=sha256:5b08fa8120f8f1b9b5658783dfee537b53cac8b295e84bfae1509d14f1b1f0da

Observation 10e67ec7-b294-463b-9d07-8daf50c0031e · outbound

This paper cites Learning to predict salient faces: A novel visual-audio saliency model,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Learning to predict salient faces: A novel visual-audio saliency model,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.665172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.410741Z digest=sha256:698832ea5067519a13e0d4b2c05c10aa995896c0da25cadafcd22a1c52d35f96

Observation 6b3a2fc4-d9c1-4727-87c3-37a87d911c38 · outbound

This paper cites Casp-net: Rethinking video saliency prediction from an audio-visual consistency perceptual perspective,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Casp-net: Rethinking video saliency prediction from an audio-visual consistency perceptual perspective,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.657150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.413446Z digest=sha256:974b1a4832d65421fe8f2ca2736439eb1b7dce0dd0b09d4ce9a2cf2a88ba0ee4

Observation b1889d04-22b4-49b3-a7e1-9227b9f3123d · outbound

This paper cites Diffsal: Joint audio and video learning for diffusion saliency prediction,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Diffsal: Joint audio and video learning for diffusion saliency prediction,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.648146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.416329Z digest=sha256:7494c7243eb73a17bbf794bd64b2ba3d708c34c15827bd7fd64289ed3db5bc54

Observation a4b3a613-a75d-409a-9d61-a63bc0025d99 · outbound

This paper cites Deep roots: Improving cnn efficiency with hierarchical filter groups,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Deep roots: Improving cnn efficiency with hierarchical filter groups,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.639371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.419165Z digest=sha256:f469b2b995e8ee8bfd9d296d7639c9f1f27288d53b786f2829ed2ff2a187d289

Observation 38379fce-13a5-4d5c-bc89-878ef2f644fb · outbound

This paper cites Shufflenet: An extremely efficient convolutional neural network for mobile devices,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Shufflenet: An extremely efficient convolutional neural network for mobile devices,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.631380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.421867Z digest=sha256:514c0ecc6772c67ddde090110aad2b28b2013861c26229ae754e1095e7c46720

Observation e9d9ee8c-c587-4518-b985-b5ab1e8436dd · outbound

This paper cites Actor- context-actor relation network for spatio-temporal action localization,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Actor- context-actor relation network for spatio-temporal action localization,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.623245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.424569Z digest=sha256:dd886c3063f8ded7332f21bda0afd0e48afc9fc010336cf3299ebb27241c3fe1

Observation 4e26c035-8412-4faa-9f0b-278113f4b862 · outbound

This paper cites Slowfast networks for video recognition,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Slowfast networks for video recognition,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.613895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.427677Z digest=sha256:16c7993d984ce21ce48c29be519481126f4f3b78098234767ff5334ae930bb8a

Observation fd2f8970-3785-4b1f-80c7-85c8171a254e · outbound

This paper cites Ava: A video dataset of spatio-temporally localized atomic visual actions,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Ava: A video dataset of spatio-temporally localized atomic visual actions,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.605955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.430271Z digest=sha256:d154e6883b5802890d5ef98f4d6f5fca058140546e1c0fcd7f53d938a313a16a

Observation 8cb2d9bb-038a-4f02-b8f1-5b6f8f2af615 · outbound

This paper cites Actions in the eye: Dynamic gaze datasets and learnt saliency models for visual recognition,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Actions in the eye: Dynamic gaze datasets and learnt saliency models for visual recognition,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.597676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.432862Z digest=sha256:609ce90cac9a870ab62082493dad2616b42e650145beed2855d0a7600ddd57e1

Observation 52e2a6d1-abd0-4b70-8af4-0adb6a543475 · outbound

This paper cites Fixation prediction through multimodal analysis,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Fixation prediction through multimodal analysis,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.589435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.435844Z digest=sha256:e6f5c88fb410fc587182a77cd20fb31d3a415cb044b829612530cff6c02b32fb

Observation 629d8538-fa1b-4492-acef-c46d43f675eb · outbound

This paper cites How saliency, faces, and sound influence gaze in dynamic social scenes,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues How saliency, faces, and sound influence gaze in dynamic social scenes,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.579363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.439170Z digest=sha256:1c69023de9ae42893b4b3ebf09b3b85d5c6f90cd056eca11b3ebf1175302ed09

Observation 0c4e7ff8-ff2c-4f33-9122-41949582ce8c · outbound

This paper cites Toward the introduction of auditory information in dynamic visual attention models,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Toward the introduction of auditory information in dynamic visual attention models,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.569696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.441720Z digest=sha256:71b53db924339c0a3f06a1ee758a26d78e296cf225225e60c6614753b512fc1b

Observation 3776c107-34d1-456c-9e00-fb98a76caced · outbound

This paper cites An efficient audiovisual saliency model to predict eye positions when looking at conversations,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues An efficient audiovisual saliency model to predict eye positions when looking at conversations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.559930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.444399Z digest=sha256:f4e25f452f67e8f29406af3075f4b5ee0e769b19ec94995cb0071637d8c1bbb7

Observation f3f96374-e6b4-4a69-b076-9742178c10c0 · outbound

This paper cites Clustering of gaze during dynamic scene viewing is predicted by motion,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Clustering of gaze during dynamic scene viewing is predicted by motion,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.549618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.447075Z digest=sha256:851904e92046b990fe035e79c43a6436a20216d9ad01a66e95d228beac10c008

Observation 3fe200b4-4cd8-4d74-8c52-66005e90d132 · outbound

This paper cites Predicting eyes’ fixa- tions in movie videos: Visual saliency experiments on a new eye- tracking database,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Predicting eyes’ fixa- tions in movie videos: Visual saliency experiments on a new eye- tracking database,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.539925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.450023Z digest=sha256:67b528f8f2c52d7fbcb275bc899bd47210c491f800c96d9293e66a4e566a97e1

Observation 4a8c7cbd-e557-4448-8782-5197ce51d1fd · outbound

This paper cites Tinyhd: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Tinyhd: Efficient video saliency prediction with heterogeneous decoders using hierarchical maps distillation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.530709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.452673Z digest=sha256:743c88251d6f0206cfed4bfd36b871c9b7afa9689f3cc9031e081b8b602dba00

Observation c0e331bd-fbe9-4858-a4ff-4d6849cba445 · outbound

This paper cites Video saliency forecasting transformer,.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Video saliency forecasting transformer,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.520968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.455473Z digest=sha256:cfdc4081175bbb1ead1d18f355e71508faec5d6d036dc3c8c0624f2d6f3ad0a6

Observation 6a52a3c4-8d4b-4839-b001-4c5fd5744ede · outbound

This paper cites What do different evaluation metrics tell us about saliency models?.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues What do different evaluation metrics tell us about saliency models?

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.512313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.458030Z digest=sha256:a4dee61c5bdb72459b61366307d3f7ecfce36e7a1e98d6d28c79158638fe1b1f

Observation 0dd95823-6a1f-4c1c-925a-c30a2d44240b · outbound

This paper cites Does audio help in deep audio-visual saliency prediction models?.

Minimalistic Video Saliency Prediction via Efficient Decoder & Spatio Temporal Action Cues Does audio help in deep audio-visual saliency prediction models?

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:13:03.501771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-09T19:13:03.461605Z digest=sha256:1ba1dbf2b761bc54b6093c4eb7c1a94e690f1af90dac8c50e8b26db8ab79070b

Pith citing papers

No inbound Pith citation observations are available.