Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T06:00:50.704572Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 0 inbound Pith citation observations for arXiv:2509.04712.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T06:00:50.704572Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee52a3a2-f5b2-4319-a44f-364d299d5565 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving A survey of autonomous driving: Common practices and emerging tech- nologies,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0cb7e546-170d-4ba8-8088-1d89d2e43943 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Lane change and merge maneuvers for connected and automated vehicles: A survey,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c93f8b65-980f-45c8-a51a-d3fb682a2c58 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Automated lane change controller design,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0f9284b2-f756-4e47-8754-778eb9f38f79 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Traffic dynam- ics: studies in car following,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f33213f1-b4a5-466d-837b-0d1dfb435c13 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving A behavioural car-following model for computer simulation,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8b341eb3-75a1-4475-bca7-c7d6930636bc · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Congested traffic states in empirical observations and microscopic simulations,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20015de6-1c5c-4f10-aa39-1b2aa8b98703 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving General lane-changing model mobil for car-following models,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6715f95c-e45f-4db3-8ba8-878f389d535c · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Driving intention recognition and lane change prediction on the highway,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c24b7b92-da87-4042-8347-13feec93eedd · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving End to End Learning for Self-Driving Cars
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 375a326f-43e9-4b16-9a36-85937ffbcd2e · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Explaining How a Deep Neural Network Trained with End-to-End Learning Steers a Car
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55b9b7e8-12ba-4701-a5b2-05f268b3e490 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving End-to-end driving via conditional imitation learning,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 36b2f5bc-7278-495e-8fb1-8b47a0162173 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Urban driving with conditional imitation learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 55972cc3-481a-460e-a29a-3e839dd90330 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9bb2ed2-0075-4f50-a6b6-bf88234ca509 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Proximal Policy Optimization Algorithms
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0beb63b8-ef8f-4cca-a6c2-26501b0a1cd7 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Soft actor- critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6c14dc3-1631-44ad-a571-876a55ac214f · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Extensive Exploration in Complex Traffic Scenarios using Hierarchical Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f30b5aa9-2645-4da6-b0cb-3c763ee0f7fa · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Exploiting hier- archy for scalable decision making in autonomous driving,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 82cfd841-8194-4b4f-a592-e118644d4a2f · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Learning hierarchical behavior and motion planning for autonomous driv- ing,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 442e7056-f63b-42d2-9cec-8e1fe35220f9 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Deep hierarchical rein- forcement learning for autonomous driving with distinct behav- iors,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9530e4af-1363-4bc5-8a88-e4d940d0f112 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving A rein- forcement learning approach to autonomous decision making of intelligent vehicles on highways,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9b9c98ce-ed79-45fd-a6dc-46550f4f3f75 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Integrating deep reinforcement learning with model-based path planners for automated driving,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3a1a448e-32d0-4c06-8cad-5fbcedbb70ff · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Lane change decision-making through deep reinforcement learning with rule- based constraints,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 202b78e6-34ee-4f55-b85d-e941086fc858 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Combining reinforcement learning with rule- based controllers for transparent and general decision-making in autonomous driving,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6948b72f-e7b5-4b3e-a656-133b06883507 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving A combined reinforcement learning and model predictive control for car- following maneuver of autonomous vehicles,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 29d92c44-477d-48f7-b5f3-3535411ff1ed · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Combining reinforcement learning with model predic- tive control for on-ramp merging,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation fb03a919-7a64-4c50-8a76-44a78b66fd6c · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving A Hierarchical Architecture for Sequential Decision-Making in Autonomous Driving using Deep Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 94b4fe8e-3b93-44d9-8ac6-7d5ae5304ca8 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Driving decision and control for automated lane change behavior based on deep reinforcement learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2ad53c79-efa3-406a-94ca-e017a3af85d7 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Prioritized experience- based reinforcement learning with human guidance for au- tonomous driving,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 27b91336-1d58-4653-afa1-78e13170692f · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Learning to drive in a day,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3b472d54-e28a-4949-b19b-de8fe36f6896 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Efficient deep reinforcement learning with imitative expert priors for autonomous driving,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6bda6f56-dce9-46bf-9fa5-8cbe3c6a03d7 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Imitation is not enough: Robustifying imitation with reinforcement learning for challenging driving scenarios,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19369424-5080-40a3-b3f4-36d24e2bd34e · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Boosted bellman residual minimization handling expert demonstrations,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 46d406f3-6250-4d75-a145-16ba275627c3 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Deep q-learning from demonstrations,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e453e3c0-12a8-4260-b396-4acb29cf25ab · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Reward learning from human preferences and demonstrations in atari,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10074819-6656-4277-8120-92136c49b0d5 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving SQIL: Imitation Learning via Reinforcement Learning with Sparse Rewards
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55434cac-0246-4a85-b7ec-c8d4ce512d71 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving An environment for autonomous driving decision- making,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 319d0981-1e94-4a81-a857-42912ccd75b6 · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Conservative q- learning for offline reinforcement learning,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation fc2ea35e-942f-4094-a434-f69b9e4cf02c · outbound
Bootstrapping Reinforcement Learning with Sub-optimal Policies for Autonomous Driving Generative adversarial imitation learning,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.