Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:41:42.393210Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 4 inbound Pith citation observations for arXiv:2508.01539.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:41:42.393210Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T11:03:51.194382Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T21:38:59.525642Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 447b2ff4-7e7c-49d7-a291-9b33bddd2fdd · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Paluch, J
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2ec8c764-df64-4d11-9d56-359dd62ce709 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a11ebb28-33ba-4a3c-93d7-9451a6be7ba2 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation BehAV: Behavioral Rule Guided Autonomy Using VLMs for Robot Navigation in Outdoor Scenes
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4176d68-2e40-4bfe-a753-303db470432b · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation ViNT: A Foundation Model for Visual Navigation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c164fd9-6367-4b9d-b349-76e0ce34874b · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation CROSS-GAiT: Cross-Attention-Based Multimodal Representation Fusion for Parametric Gait Adaptation in Complex Terrains
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 008c4d89-bd89-41cb-aa6b-87abc623945c · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Social-LLaVA: Enhancing Robot Navigation through Human-Language Reasoning in Social Spaces
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26bce7b1-d0f1-4233-9dc8-683c84c58121 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Offline Reinforcement Learning for Visual Navigation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 211d81ce-997e-4b8c-81cd-243c1b8c3166 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 84047cca-00d1-41af-b27e-2ba892b5c6f4 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Improving Generalization in Reinforcement Learning Training Regimes for Social Robot Navigation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cd5aadd-c4bd-4ac0-988b-5d5c93c4c7e1 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation SoNIC: Safe Social Navigation with Adaptive Conformal Inference and Constrained Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce30310a-c7da-46ee-824a-a33cf9d33c36 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Jiang, P
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d580991e-99c6-491a-a5a2-e269b95bd9f0 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Karnan, A
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d6cab99-410d-44ea-bd41-66f6294efe59 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation VAPOR: Legged Robot Navigation in Outdoor Vegetation Using Offline Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 769af3b3-c51b-4e79-9bf4-39931cc1d620 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Caesar, V
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dcbe0c9-1d50-4df9-ac46-93983f6cb10e · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation A2D2: Audi Autonomous Driving Dataset
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62aec65-6ef1-4b97-a5af-745d3c83d26e · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Kumar, A
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 514ea7c8-6cc3-4cf4-a0ea-3e4959d22580 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Kapoor, S
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f417724-f7b0-4cad-82a6-103914ca6f3d · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Liang, U
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8aa917d6-18e0-4a3d-a73e-dfc105c476ac · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Patel, N
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a9029285-c2f9-4d4c-82fa-3c3c26941291 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Beyond Accuracy: Evaluating the Reasoning Behavior of Large Language Models -- A Survey
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db93ab8e-02c2-44f8-960c-e07eed811093 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Towards Reasoning in Large Language Models: A Survey
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5dbacfa-40e5-4177-9235-3f517a90ed75 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d3615b1-23d6-47be-8485-1f2fc8897293 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73adf8aa-a8c5-42b2-8ad9-bc7b26c7ca4f · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Huang, O
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 972c6a70-f895-47ea-9862-77cb281e1147 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 11d7e473-5c36-4862-937d-a2013cca1700 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2895e9fc-018e-4b83-9b35-57610a6036af · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cf4ee867-a2ce-44d8-81f7-cb2626279152 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Resilient Timed Elastic Band Planner for Collision-Free Navigation in Unknown Environments
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2e6bb728-0397-46b0-842a-74c74711489b · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Dosovitskiy, G
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd0c8b2-bb76-49eb-9889-9f587c83d6ab · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c991c3d9-3363-4ab2-af98-b60d1ebf6e51 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf990ad4-6fe6-4e04-848f-c4f42f575879 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1debe993-d244-4a08-84cc-9bc6ccd864dc · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8be3d71-d44e-4dd6-a8b7-807fd669d6ff · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4ff4c6b-9d2e-4e44-9182-ff5299d37b8d · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9187e42d-90ed-4e90-87c8-1db10a76ff20 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0c29b2-a42e-48d3-a0f6-0bab735b5f1c · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation VL-TGS: Trajectory Generation and Selection using Vision Language Models in Mapless Outdoor Environments
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 889620ba-285b-4180-a4bd-b176174195c5 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bc873a4-4dbe-4ee1-8c79-250333343146 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4639882f-adba-4a59-be20-18e10c678589 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Oquab, T
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 188869df-076a-41a1-97c4-638f36f7f3db · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Offline Reinforcement Learning with Implicit Q-Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49c3d291-48b3-475d-a818-76b8f1702a53 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Fujimoto, H
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d25c821-87d7-4992-9060-40219e793cba · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb7ede81-6f6f-4105-900d-e51358250c5b · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Nazeri, J
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 379f94f0-fa1b-4266-96ec-9aba6588e650 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Alt and M
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 84f69c58-bd53-4431-889f-5e35c3b68a4b · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Fujimoto and S
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b3622393-bca0-4325-b54e-97a502504c47 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Tarasov, A
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7deac5a7-971a-4e7e-9fcb-c0f8a8508c9e · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 82fe68fc-0203-4da4-8ec9-cc343b14cfca · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6cda450d-7312-4cda-8761-650da2fc7941 · outbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eee1caca-fdb9-4b91-92f1-fe5dd3906958 · inbound
Interpreting Context-Aware Human Preferences for Multi-Objective Robot Navigation HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c3309775-2f63-4448-8057-d8ba1c33babe · inbound
From Imitation to Alignment: Human-Preference Flow Policies for Long-Horizon Sidewalk Navigation HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 573bdfba-b5b0-44f2-ac81-57da7299b166 · inbound
VEGA: Learning Navigation VLAs from In-the-Wild Egocentric Video with Geometric Trajectory Supervision HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a1802dc2-07e2-429d-8bc2-5355b01cac3f · inbound
VEGA: Learning Navigation VLAs from In-the-Wild Egocentric Video with Geometric Trajectory Supervision HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.