Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T17:13:04.236686Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2606.09142.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T17:13:04.236686Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b3b16424-1d11-4892-a331-214cc0b4fcae · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Pedestrian Behavior Prediction Using Deep Learning Methods for Urban Scenarios: A Review,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bed714ee-335a-474a-9ccf-df0bec14183d · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Predicting Pedestrian Crossing Intention in Autonomous Vehicles: A Review,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 383f2029-8a44-4880-9512-a11756f0e28d · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models EgoNav: Egocentric Scene-aware Human Trajectory Prediction
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07d7f1f7-3e63-4e8b-b169-94183c628fef · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea260b1b-b7f3-4009-b230-f2259288b332 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Ego4d: Around the world in 3,000 hours of egocentric video,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4726cbaa-423f-4dfa-bd47-047b4cc63bcb · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models An outlook into the future of egocentric vision,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e86667fc-7de0-4676-9c80-3a6280931162 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Egolife: Towards egocentric life assistant,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 045b6d09-7e66-4972-bcb1-c3fa6acd7e14 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models EgoCogNav: Cognition-aware Human Egocentric Navigation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6facd0e-cec0-4beb-b701-c4b755743cf4 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Lookout: Real-world humanoid egocentric navigation,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8e202bb-8cc7-4911-81f0-d38c8fe1147c · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models HEADS-UP: Head-Mounted Egocentric Dataset for Trajectory Prediction in Blind Assistance Systems
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb1ebd8b-00a6-4617-8d11-29e556b205ee · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Egocentric human trajectory forecasting with a wearable camera and multi-modal fusion,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85033b4b-6be4-41f6-99b7-c197bf8f19e4 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models KrishnaCam: Using a longitudinal, single-person, egocentric dataset for scene understanding tasks,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 833541b9-b19d-4a4c-8302-bd06132b7b72 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Egocentric future localization,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1078c9ee-9460-4c66-845d-30830139088d · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Pedestrian intention prediction for autonomous vehicles: A comprehensive survey,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0915f032-7cf4-4331-a100-210cd1ed0c95 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models GPT-4V Takes the Wheel: Promises and Challenges for Pedestrian Behavior Prediction,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d64df1dd-9b5c-48ef-81d1-53719e388f7b · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models OmniPredict: GPT-4o Enhanced Multi-modal Pedestrian Crossing Intention Prediction,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 451f2793-aea3-46ec-a694-3985bab79be0 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Seeing beyond frames: Zero-shot pedestrian intention prediction with raw temporal video and multimodal cues,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3a739d3-43fe-47ca-83a2-37f147ebb3be · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Pedestrian Intention Prediction via Vision-Language Foundation Models,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba4a1df2-e453-46fe-a49b-631b7d2dba3b · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Optimizing Vision-Language Model for Road Crossing Intention Estimation,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2c2e360-22a0-4e17-8e67-e9159ec078d0 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Pedestrian Vision Language Model for Intentions Prediction,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aee234ee-929e-4b32-8ae2-c61d2a8c9c63 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Application of Vision-Language Model to Pedestrians Behavior and Scene Understanding in Autonomous Driving
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30a7d0f3-9a98-4bf9-84c2-d60d19d40505 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Vlmped-cot: A large vision-language model with chain-of-thought mechanism for pedestrian crossing intention prediction,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f78746e-079d-4487-a15a-6bf6ffe3da56 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Scaling Egocentric Vision: The EPIC-KITCHENS Dataset
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d07f28ea-020a-470b-a181-e3513eaab16b · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Actor and observer: Joint modeling of first and third-person videos,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6876619-31bf-4cd8-9876-0942c10bed97 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Egovlpv2: Egocentric video-language pre-training with fusion in the backbone,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 238674b4-3214-4c28-90ac-257b3e17703c · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08552ab6-46d6-4623-be20-038fdb3328b9 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Video question answering: Datasets, algorithms and challenges,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84b83758-0e70-4074-8500-882c4cf276fa · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Video question answering via gradually refined attention over ap- pearance and motion,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 526e1b70-464a-40e5-b994-342bd2f59207 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Tgif-qa: Toward spatio- temporal reasoning in visual question answering,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46884596-2fe6-4aca-bb42-45f976ef3b8e · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Activitynet-qa: A dataset for understanding complex web videos via question answering,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d53d2105-5713-4960-a549-09c8bdf2b330 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Next-qa: Next phase of question-answering to explaining temporal actions,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11f21bec-81d8-46b3-a156-6b2ddc5f1c30 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Agqa: A benchmark for compositional spatio-temporal reasoning,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0583cc37-0857-4cdb-9046-52b96999dc56 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models From representation to reasoning: To- wards both evidence and commonsense reasoning for video question- answering,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b5648aa-5505-41a4-b356-09ca8883ed37 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Egotaskqa: Understanding human tasks in egocentric videos,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 394094fd-5615-476e-bf62-a9a7fce3d8fe · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Intentqa: Context-aware video intent reasoning,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dab975f-bb5c-49d6-8f2a-dbbb3f5f33b4 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models In the eye of mllm: Benchmarking egocentric video intent understanding with gaze-guided prompting
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 531b9dbd-e5fa-4fc7-9901-f87aa630568e · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Qwen3 Technical Report
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73064eea-3446-4818-a45d-46df9badeba7 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 791ddcdc-faf3-4be6-b06e-8094bc62c4c1 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Grounded question-answering in long egocentric videos,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6007c9d7-e2e1-48fc-8187-659d98fae112 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Chain-of-thought prompting elicits reasoning in large language models,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5875e811-4bf9-452d-91b9-2b0981eff465 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f58249e3-3ca7-4fa8-84bf-4296dfe282d1 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Fine-grained visual prompting,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 384b6566-a00f-402e-abaf-f675896a50aa · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Large language models are zero-shot reasoners,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb17fc77-4a29-4ce5-b5d5-a2ca876ba246 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Grounding dino: Marrying dino with grounded pre-training for open-set object detection,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 665f8116-a965-402a-8e22-270f8ac8d387 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Simple online and realtime tracking,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39d58828-2788-459b-b4a4-2623688d79b3 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Analyzing the behaviors of pedestrians and cyclists in interactions with autonomous systems using controlled experiments: A literature review,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fb200aa-dc60-4f49-a0d5-c456b8f9001d · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Challenges and trends in egocentric vision: A survey,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbb3c8c5-db49-4ea4-98f1-cfa1e63911b3 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Eye Gaze-Informed and Context-Aware Pedestrian Trajectory Prediction in Shared Spaces with Automated Shuttles: A Virtual Reality Study
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6a9dcb69-82a4-4c93-a649-fda35ff0141c · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Lora: Low-rank adaptation of large language models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd29ecfd-42f1-46db-84c6-7f9c799a6921 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Learning Transferable Visual Models From Natural Language Su- pervision,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a80980e7-0192-4524-a0ef-2b0480689392 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Advancing Egocentric Video Question Answering with Multimodal Large Language Models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 28a30740-3f42-43dd-bc95-8acb2038985a · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9494379a-d940-41a3-9c29-07a85af5f911 · outbound
Decoding Pedestrian Crossing Intention from Egocentric Vision via Vision Language Models Chrono: A simple blueprint for representing time in mllms,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.