Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:59:56.791343Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 1 inbound Pith citation observation for arXiv:2507.04141.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:59:56.791343Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T13:25:04.053283Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T13:34:40.932917Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 89755a4e-910f-4fd8-b046-ae3d7a248e5b · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Autonomous vehicles that interact with pedestrians: A survey of theory and practice,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cc181ab-3fb6-4ce6-86b4-dfc0be51a1ff · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pedestrian intention prediction: A convolutional bottom-up multi-task approach,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64c0a2e8-e139-4b94-9995-dd787f946b28 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models St crossingpose: A spatial- temporal graph convolutional network for skeleton-based pedestrian crossing intention prediction,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8a87e33a-a222-4467-b198-55377e59f960 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models PIP-Net: Pedestrian Intention Prediction in the Wild
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0757d55c-1420-4cf7-b0c6-b743acbad52b · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pedestrian action an- ticipation using contextual feature fusion in stacked rnns,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fb433de-cfe6-4c8a-aea3-b8e46effe73e · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Do they want to cross? understanding pedestrian intention for behavior prediction,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 096beed4-6d41-4232-b594-91f5044818fc · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Multi-modal hybrid architecture for pedestrian action prediction,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 57272dc5-5f20-416e-9830-2ae0172d7b8e · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pedestrian graph+: A fast pedestrian crossing prediction model based on graph convo- lutional networks,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d431e29-b0a2-484d-924a-97647a9e5bcc · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Visual reasoning using graph con- volutional networks for predicting pedestrian crossing intention,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4e1f283-16ef-4438-84c3-4698cef4f6e0 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models CAPformer: Pedestrian crossing action prediction using transformer,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7521fcd7-f346-42b0-82ff-37c9cd83ea1b · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pit: Progressive interaction transformer for pedestrian crossing intention prediction,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 841f712e-7dfd-4388-9ebb-c567f59da6c3 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Predicting pedestrian inten- tions with multimodal intentformer: A co-learning approach,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbc0cc2c-ff16-477a-aa65-247a29a4ad34 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pedestrian behavior inter- pretation from pose estimation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7e25164-28e8-4cf8-9723-5db6a1c92a99 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Multi-scale pedestrian intent prediction using 3d joint information as spatio-temporal representation,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d72d0405-b7da-412a-ab62-1046703f990a · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Spatiotemporal relationship reasoning for pedestrian intent prediction,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42de927b-2ff4-4089-ad57-1ffa7ba398ac · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Real-time intent prediction of pedestrians for autonomous ground vehicles via spatio-temporal densenet,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aa461dfd-b696-4dcc-9b73-896c835a4f53 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pedestrian-vehicle information modulation for pedestrian crossing intention prediction,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 615b7547-8719-485d-bfb7-e2ac534abdcf · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Causal reasoning in typical computer vision tasks,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18069a05-b99a-4246-9453-288c0dc92147 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Vision language models in autonomous driving: A survey and outlook,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35d3e16b-b52f-4d17-b32c-0190c0006e9a · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Gpt-4v takes the wheel: Promises and challenges for pedestrian behavior prediction,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef2988a4-73fc-4e49-8be5-1a12e03fc931 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Omnipredict: Gpt-4o enhanced multi-modal pedestrian crossing intention prediction
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a05142d2-747f-4bde-8387-19af28ae0283 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Pedvlm: Pedestrian vision language model for intentions prediction,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf34fca6-b13d-4b6f-8afd-f1890ef41e1f · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Cross or wait? predicting pedestrian interaction outcomes at unsignalized crossings,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5fdb116b-8981-422a-8255-343c2ee44d16 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Feature Importance in Pedestrian Intention Prediction: A Context-Aware Review
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ca7ff6da-26b8-4be8-b4bd-6873e3b2a77b · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Hierarchical Prompting Taxonomy: A Universal Evaluation Framework for Large Language Models Aligned with Human Cognitive Principles
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc62114f-4913-4812-9b21-2bc71f52489a · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Large Language Models Are Human-Level Prompt Engineers
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ceed48e-7f92-4f3b-9b9f-8c3bd66adac8 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Benchmark for evaluating pedestrian action prediction,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99d7b5b0-b85f-469a-8584-a08fa55707f7 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Better Zero-Shot Reasoning with Role-Play Prompting
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 699aaa9e-581e-4a0c-afd9-1752e25544cb · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cd8d6af-bea9-49c0-8646-1cf809825b86 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Is the pedestrian going to cross? answering by 2d pose estimation,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4e747a18-0a1f-454e-8ab8-2bf16ea658da · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Chatgpt: Generative pre-trained transformer,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35420cca-34d7-4473-94d1-ff1a8173274b · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models Agreeing to cross: How drivers and pedestrians communicate,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5cf1eb8-ead7-4137-abbd-bfce2e73541e · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models PIE: A large-scale dataset and models for pedestrian intention estimation and trajectory prediction,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 398caa99-291c-4851-bb0a-8b63ce9b5728 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models GPT-4 Technical Report
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 768bfb91-1616-4cd3-82f5-8b553a37cf52 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f989038e-d8a4-4f79-a1a3-12bf90ba02b3 · outbound
Pedestrian Intention Prediction via Vision-Language Foundation Models LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b661ac7-40da-44c8-b177-06d0579874c9 · inbound
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction Pedestrian Intention Prediction via Vision-Language Foundation Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.