Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:25:27.472156Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 4 inbound Pith citation observations for arXiv:2412.13187.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:25:27.472156Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:04:30.302310Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T22:32:11.045924Z
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0d3e6144-e524-4436-8027-aac64176c2e6 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6346ffee-166d-4ce6-ab59-9127fedfa295 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Black, Danica Kragic, and Hedvig Kjellstr ¨om
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9cce5b6c-6412-479e-9e93-5e5e52b30d7a · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction We also conduct an ablation study on the zero-shot chain-of-thought (Wei et al., 2022; Kojima et al.,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e7e6acfc-e4dd-421f-9f8b-6b9e7beb43ac · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Sanja Fidler, Antonino Furnari, Evangelos Kazakos, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, et al
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 92c1b10a-8acc-4c17-9bce-f7faafe0f1b7 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Learning a hierarchy of discriminative space-time neigh- borhood features for human action recognition
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 076257f0-de31-4348-a465-e026704f16b8 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Forecasting human-object interaction: joint prediction of motor attention and actions in first person video
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 96832fac-19f8-44d5-bfb5-4f9399d95154 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca1d9f18-2315-4fd7-abbe-e261016efe07 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction R3M: A Universal Visual Representation for Robot Manipulation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfcc605b-ea90-4a26-8f5f-014e99cc8098 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction FrankMocap: Fast Monocular 3D Hand and Body Motion Capture by Regression and Integration
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0535b06f-adc5-45a9-9480-c72ded5d9914 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a71a723-d1f5-46f7-8e1c-f8af06a7f56f · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction LISA++: An Improved Baseline for Reasoning Segmentation with Large Language Model
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1de86b4-749c-4e62-aea8-a3a61477779d · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a89cbccd-29c9-442b-bd8e-955d67dff461 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61021fcc-290b-4904-9116-01d79faf9edf · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Pix2seq: A Language Modeling Framework for Object Detection
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a9f8887-2b0d-4364-9118-47c4967f1d40 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Judith B¨utepage, Hedvig Kjellstr¨om, and Danica Kragic
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation dfde9f0b-b4d6-455d-9079-bf0a30581932 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Unresolved cited work
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60e9aecc-25c2-475e-84be-71171ad4ec03 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction LITA: Language Instructed Temporal-Localization Assistant
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 621224e9-6890-40f9-94c7-cf596e8ae6f8 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e7e6801-9363-4d16-9c09-aa326ae7d8b2 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Expressive Forecasting of 3D Whole-body Human Motions
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6413f722-0f89-42c6-98a4-9dc500c3ce26 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction Uncertainty-aware State Space Transformer for Egocentric 3D Hand Trajectory Forecasting
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b78afefe-1b21-4649-9f04-ad02675583b9 · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction OpenVLA: An Open-Source Vision-Language-Action Model
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 635f13ad-a235-49ee-8cee-bfc845328e1a · outbound
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3e68d82-77d7-40aa-8952-8f19c7507f26 · inbound
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cea29c6-c471-4e35-83b3-4124dd3243cf · inbound
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 35dbf9cc-5449-4140-a493-34be081ab8b5 · inbound
EggHand: A Multimodal Foundation Model for Egocentric Hand Pose Forecasting HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 89482c55-4d31-4fb9-8ee6-08a30eaafcdd · inbound
MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.