Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:11:46.728553Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2506.08649.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:11:46.728553Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d549bf15-a081-4f71-99d9-3c23b26c444d · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Beit: Bert pre-training of image transformers, in: International Conference on Learning Representations, pp
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d4533415-1463-4bb9-ade0-ebc2debd5557 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Emerging properties in self-supervised vision transformers, in: Proceedings of the IEEE/CVF international confer- ence on computer vision, pp
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5f6d40fb-6dbc-4058-ade9-bc74f41d3314 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Quo vadis, action recognition? a new model and the kinetics dataset, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 781eb04d-9054-47f5-b515-80bf6f8f6ffd · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization An empirical study of training self-supervisedvisiontransformers,in:ProceedingsoftheIEEE/CVF International Conference on Computer Vision, pp
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f7f2fcf9-4a59-4e6d-84ee-f354477db27b · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Videomem:Constructing,analyzing,predictingshort-termandlong- term video memorability, in: Proceedings of the IEEE/CVF Interna- tional Conference on Computer Vision, pp
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5073d918-1ea4-4e6f-9efb-10efdc8bca9f · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Annotating, understanding, and predicting long-term video memora- bility, in: Proceedings of the 2018 ACM on International Conference on Multimedia Retrieval, pp
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6095a6a9-f846-4065-a195-77aa90910f05 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ffd4819-98e8-4c73-b6fe-5d4919e5d333 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Aimultimedialab at mediaeval 2022:Predictingmediamemorabilityusingvideovisiontransformers and augmented memorable moments , 12–16
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ce3ed93-6776-41c0-901a-45c0efaa3776 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e36c3a1-ad21-47ba-8486-1d1bdfc0bd25 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization An image is worth 16x16 words: Transformers for image recognition at scale, in: International Conference on Learning Representations, pp
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a9e927e-a2ce-4c9b-8ee3-b13738bf9dc4 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Modular memorability: Tieredrepresentationsforvideomemorabilityprediction,in:Proceed- ings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 84ec52ee-4516-4af3-b617-78011458f673 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Memory: A contribution to experimental psychology
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de1c3754-e5bd-4c67-8476-58f973fd7335 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Amnet: Memorability estimation with attention, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0f6c0f8-f23b-4f2e-94c7-236cc1b7868c · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Supervised video summarization via multiple feature sets with parallel attention, in: 2021 IEEE International Conference on Multimedia and Expo, pp
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation be5b7807-0dae-4325-b6c1-3758efa4560c · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Creating summaries from user videos, in: Computer Vision–ECCV 2014: 13th European Conference, pp
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd0c0786-15b0-4c1a-921d-bab3f1d12ecb · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Learning computational models of video memorability from fmri brain imag- ing
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ffe99691-de32-4b92-ba29-32342f3291fb · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Self-supervised co-training for video representation learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ede9dd67-556e-44a8-a51b-f425d216a1c5 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Can spatiotemporal 3d cnns retrace the history of 2d cnns and imagenet?, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0174df5-c239-4170-b3ba-5ff567d05f2d · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Momentum contrast for unsupervised visual representation learning, in: Proceed- ings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a059f577-50fa-40ae-9079-d16b7f94146b · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Deep residual learning for image recognition, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dec3552-4bd3-42df-a5d2-3c1405283945 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Densely connected convolutional networks, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce92ed41-a7c0-464d-836a-ce8fd1d80404 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Whatmakesanimage memorable?, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2ea92ec0-1555-4f97-9453-a060091b111f · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bea60b69-7984-4d34-b94b-86b41c1fe728 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Understanding and predicting image memorability at a large scale, in: Proceedings oftheIEEEInternationalConferenceonComputerVision,pp.2390– 2398
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91762437-a41f-4ea4-89f3-4819249ab71e · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Topic-oriented text features can match visual deep models of video memorability
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72896d0d-303f-47a7-862a-0fd231a61be8 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 48757189-2542-41bc-aac0-c056b1ca5b4b · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Scene memory is more detailed than you think: The role of categories in visual long-term memory
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f08918d2-d602-4c82-9f48-86c00bf63e79 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Multimodal deep features fusion for video memorability prediction, in: Working Notes Proceedings of the MediaEval 2019 Workshop (CEUR Workshop Proceedings), pp
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e034b08a-7b4c-4fec-8f8a-0b6c0e015bf6 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Adaptive multi- modalensemblenetworkforvideomemorabilityprediction
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71079d0d-67ba-4b7e-b32f-27c3ec9ed57d · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Deephierarchicallstmnetworks with attention for video summarization
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff956c57-9814-42bf-b4dd-ec3d6c424eae · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 23248fe8-ccad-4499-a197-a514ea451e42 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91ecfd1f-9f8f-41fa-a58d-052615cc9d23 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Video storytelling based on gated video memorability filtering
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1d0b8dae-449e-4cbc-aff0-1aaf2fabcbce · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Audio-visual instance discrimination with cross-modal agreement, in: Proceedings of the IEEE/CVFConferenceonComputerVisionandPatternRecognition, pp
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97421f1b-278c-400a-9857-ab4a7a232d88 · outbound
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 570c3d7d-ef52-4e09-8b49-7ed65ff3e591 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Multimodal memorability: Modeling effects of semantics anddecayonvideomemorability,in:ComputerVision–ECCV2020: 16th European Conference, pp
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8906cfb4-8275-4239-b9c0-d7363c616322 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Spatiotemporal contrastive video representation learning, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47a9df1e-e41e-4f9d-b905-a0743f5d9dff · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Does Video Summarization Require Videos? Quantifying the Effectiveness of Language in Video Summarization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50758686-3c06-448b-8f17-18b60185de27 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Self-supervised video transformer, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6bba91b0-db5e-4035-b671-d0e3647c7f9b · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Ex- ploringmultimodality,perplexityandexplainabilityformemorability prediction, in: Working Notes Proceedings of the MediaEval 2021 Workshop (CEUR Workshop Proceedings), pp
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 326a40c3-738d-40ba-a555-8067694ee21c · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Learning transferable visual models from natural language supervision, in: International Conference on Machine Learning, pp
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5c39147e-8b28-4b8b-9caa-19a1c1050184 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2f732db-0da7-458a-8223-282ff3392ce8 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization A network linking scene perception and spatial memory systems in posterior cerebral cortex
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation afbe63c6-0be1-47be-af86-05dae2b3d1fe · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Tvsum: Summarizing web videos using titles, in: Proceedings of the IEEE conference on computer vision and pattern recognition, pp
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d06d11f4-aae0-4123-9feb-a02475e4e491 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Predicting media memorability:Comparingvisual,textualandauditoryfeatures,103– 105
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 929fe93a-ef38-4ba4-8e2b-b044a5a33f5e · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Diffusing Surrogate Dreams of Video Scenes to Predict Video Memorability
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c5972c9-04a2-4f90-b2b6-8b390cfe2712 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Siamese image modeling for self-supervised vision representationlearning,in:ProceedingsoftheIEEE/CVFConference on Computer Vision and Pattern Recognition, pp
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88fcf709-254d-43dc-94db-1cbe309cd1de · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Recurrent unit augmented memory network for video summarisation
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3c5dd633-7268-40ca-b90e-a4cbc551ed13 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Modelling of video memorability using ensemble learning and transformers , 7–11
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f3b7039-30bb-47c2-9dd0-42c8088d0041 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Reconstructivesequence-graph network for video summarization
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47d8ed23-a8be-460a-9a4c-4d5c6f4ef6b5 · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Learning multiscale hierarchical attention for video summarization
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f87002f-97f5-4cc9-a326-97089498140c · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization Training data-efficient image transformers & distillation throughattention,in:InternationalConferenceonMachineLearning, PMLR
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f02226-52a6-4ebb-824f-cb86936bc0d2 · outbound
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7cf91807-10fb-4bf6-88e2-4e356758669b · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization IEEE Transactions on Pattern Analysis and Machine Intelligence 44, 4065–4080
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9de527e-2891-4ccf-a400-2d771e263e9a · outbound
Enhancing Video Memorability Prediction with Text-Motion Cross-modal Contrastive Loss and Its Application in Video Summarization IEEE Transactions on Image Processing 31, 1573–1586
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.