Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:23:20.545675Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2501.05057.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:23:20.545675Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:13:41.473724Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T15:13:45.789156Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9c0ec740-bba7-4405-8a47-d038f30db641 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A Survey of Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f65d00d-0e1e-4896-8256-b0fe8a12b4c6 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A survey on evaluation of large language models,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae648a41-b746-4658-b9ed-2f1b030fcbcf · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A survey on multimodal large language models for autonomous driving,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a7e49d22-0725-4ac4-8da7-e1255f1f20db · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Deep reinforcement learning for autonomous driving: A survey,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea9d91e9-c1f6-44ad-b7ef-964d33b18b0f · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Deep learning for safe autonomous driving: Current chal- lenges and future directions,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 00b6254c-2a1f-4dad-bdc7-c479e461d9ca · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Deep learning-based vehicle behavior prediction for autonomous driving applications: A review,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5ea2442-f551-4468-b264-1789558119d7 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Large scale interactive motion forecasting for autonomous driving: The waymo open motion dataset,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e3956e3-dd62-4281-bbfe-c1c5f53e2d80 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Behavior planning at urban in- tersections through hierarchical reinforcement learning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 15a9f599-75a0-4cc7-b189-f148df4aa625 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Chance-aware lane change with high-level model predictive control through curriculum reinforcement learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0971f943-c18e-4f8d-9024-6027493be1e2 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Reward-driven automated curriculum learning for interaction-aware self-driving at unsignalized intersections,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e9444779-07ff-435f-a925-3f364801783c · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models The perils of trial-and-error reward design: misdesign through overfitting and invalid task specifications,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5da25052-e7f1-4be3-a9ea-93126015eb4e · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Learning to utilize shaping rewards: A new approach of reward shaping,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d40221e8-5493-491d-98e3-ca6f40d631bb · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cac0da9a-ac47-4e7d-8d83-803a0338381f · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d2f8eab-d19b-45b9-9233-9db41527a231 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A review of reward functions for reinforcement learning in the context of autonomous driving,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c157aaa0-3b1c-412d-8874-db0dea0c39b8 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum learning for reinforcement learning domains: A framework and survey,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cf542e1-e1ce-49ae-8b95-36c87579ab12 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum learning,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b4a500-4286-494b-af53-d1f9019821ea · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum learning: A survey,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffdddfd3-5628-497a-9d02-15dfb15d5822 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Automated curriculum learning for neural networks,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f02328a6-984c-4730-a035-0429e9ec7ae9 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Robot parkour learning,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 370b7dc6-e493-4a48-8431-264a9d48ecc8 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Self-learned autonomous driving at unsignalized intersec- tions: A hierarchical reinforced learning approach for feasible decision- making,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a7db27e1-0d20-4fe9-8814-b205fbb474bd · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A multi-task reinforcement learning approach for navigating unsignalized intersec- tions,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f1524a2a-24fb-40c8-9c58-4ca94b1c053e · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Multi-task reinforcement learning with context-based representations,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 33c73be2-f491-47d7-9848-65a69a03f3c0 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Multi-task safe reinforcement learning for navigating intersections in dense traffic,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6b6c8c8-a115-4361-88db-723a0131e68a · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Learning robust rewards with adver- sarial inverse reinforcement learning,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 789bb713-f49d-43d3-b83a-ad2bf136fa4e · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A survey of inverse reinforcement learning: Challenges, methods and progress,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 28f613bb-6432-4abc-a2c8-3a3a042ffbf2 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Driving in Real Life with Inverse Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b41c043b-1c5e-4adc-a821-4921d35ed981 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Interaction- aware planning with deep inverse reinforcement learning for human- like autonomous driving in merge scenarios,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 789b3bd2-a072-4580-9df5-9c3794684af9 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Exploration-guided reward shaping for reinforcement learning under sparse rewards,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f299bd2-c4e6-4bba-8554-bc866c1063ab · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Reward design with language models,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 446c984c-96db-47ba-b3e0-1cd406111cdf · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Auto mc-reward: Automated dense reward design with large language models for minecraft,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 59b10c27-b8e1-485b-910e-f3987c1abe18 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Centralized cooperation for connected and automated vehicles at intersections by proximal policy optimization,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c716c27f-ea00-4d90-869d-73f577d50742 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Au- tonomous overtaking in Gran Turismo sport using curriculum rein- forcement learning,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9cf8110e-3dcd-4e19-946f-b4a48bf8ecf8 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum proximal policy optimization with stage-decaying clipping for self- driving at unsignalized intersections,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e1aee4d8-f551-4c9a-8e0f-c30ce08419c8 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Automatically generated curriculum based reinforcement learning for autonomous vehicles in urban environment,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0416e0e1-a419-4132-bc30-1490c7312021 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models State dropout-based curriculum reinforce- ment learning for self-driving at unsignalized intersections,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cdb15ce1-6d61-499a-8a4d-533f928856c0 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Dilu: A knowledge-driven approach to autonomous driving with large language models,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e259233e-f6b9-4eec-900a-137ebd7b56ee · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Drivegpt4: Interpretable end-to-end autonomous driving via large language model,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96792b4e-62bb-4162-aef5-c785c1a88822 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2754f77-3c41-4a8c-bb7b-857ce9da54cd · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02c69e66-2a6d-44f3-a89e-e2256b3ab263 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Language to 14 rewards for robotic skill synthesis,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 68c4d726-39aa-4b74-bb02-d46ad8466384 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models REvolve: Reward Evolution with Large Language Models using Human Feedback
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddcb794a-3c12-4229-af19-995f8f9f066a · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Human-centric Reward Optimization for Reinforcement Learning-based Automated Driving using Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffc041d8-05e7-4446-bb53-c5b3fd2806a1 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Dreureka: Language model guided sim-to-real transfer,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5581ed92-4026-4e92-9517-87ced3da2592 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 578f496c-fa07-438b-99d2-7b878d2c17eb · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Autoreward: Closed-loop reward design with large language models for autonomous driving,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5daacbbd-efdb-4c6d-8b54-aad17f5a9e0d · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models CARLA: An open urban driving simulator,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b92082a0-f68c-421f-b25e-6141fb0a4740 · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Proximal Policy Optimization Algorithms
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 050329e2-104a-4e2a-a88b-f2352da2604c · outbound
LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Casadi: a software framework for nonlinear optimization and optimal control,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eedf6962-2325-4712-84a7-9fae99fe9538 · inbound
HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.