Pith. sign in

Paper Citation Record · LEDGER

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.16336.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16336 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:14.797853Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact7
  • verified fuzzy11
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 117b6a38-8987-42aa-815d-6ed68ac816d6 · outbound

This paper cites Autonomous Driving at Unsignalized Intersections: A Review of Decision-Making Challenges and Reinforcement Learning-Based Solutions.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Autonomous Driving at Unsignalized Intersections: A Review of Decision-Making Challenges and Reinforcement Learning-Based Solutions

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:16.855488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:09.979709Z digest=sha256:a72c1000d4a3feb287300ef74215ef0575787e36cb6f0cf83911aca6d5339615

Observation 0ee6d5c9-79b8-4efb-9c2a-ac0ea1a8ea53 · outbound

This paper cites Enhanced intelligent driver model to access the impact of driving strategies on traffic capacity,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Enhanced intelligent driver model to access the impact of driving strategies on traffic capacity,

Reference 2

Resolution
verified exact
raw_fallback, observed 2026-08-06T23:49:16.567981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:10.045389Z digest=sha256:48baeab49ad2145df7d33b10b91baefb2af4a33d3f5e32f92f074ca8446abfd2

Observation ab3dbb70-193c-45c9-bf85-6051c9ce7b9f · outbound

This paper cites Extended safety descriptor measurements for relative safety assessment in mixed road traffic,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Extended safety descriptor measurements for relative safety assessment in mixed road traffic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:19.466504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:10.208393Z digest=sha256:bdf633cf051227727e5ca92d4ba82b64e075e7d8f4f85fb1724518143b677e1f

Observation 26bb3be2-7e4b-497b-84b4-8774b67b1c5e · outbound

This paper cites Analysis of optimal velocity model with explicit delay,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Analysis of optimal velocity model with explicit delay,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:10.359278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:10.359278Z digest=sha256:c8a265a48ccebc263b4e3aa9fe503578d18b418ecb37708391ad07e88e81bbe4

Observation f4ed6449-325b-4c5a-9c88-29689429fb15 · outbound

This paper cites End to end learning for self-driving cars,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections End to end learning for self-driving cars,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:19.321928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:10.458971Z digest=sha256:07ec5582d8ae31b362b6e4c64c920126f7e4f728cf63857af939969a37ea8558

Observation 5f3ccb8d-1c36-4a81-a3fb-29b51c65009f · outbound

This paper cites ChauffeurNet: Learning to Drive by Imitating the Best and Synthesizing the Worst.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections ChauffeurNet: Learning to Drive by Imitating the Best and Synthesizing the Worst

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:10.703645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:10.703645Z digest=sha256:003e66f9f2fa33ac6876ab78fcbc812342e46de727b094091e165e03d8e8b2d4

Observation df52118b-0f50-4af1-b9f2-511523201f97 · outbound

This paper cites Navigating Occluded Intersections with Autonomous Vehicles using Deep Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Navigating Occluded Intersections with Autonomous Vehicles using Deep Reinforcement Learning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:16.254374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:10.825859Z digest=sha256:8ccdea7d7660d0aab7271c489e096fcf7dfd9820d857a5824ac0a65d18867a85

Observation 41825fd5-c11b-4f55-b155-c0c9ee38ffed · outbound

This paper cites Learning Negotiating Behavior Between Cars in Intersections using Deep Q-Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Negotiating Behavior Between Cars in Intersections using Deep Q-Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:16.032798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:10.964587Z digest=sha256:34eb3ee0f23f75ebbf240f510711bb82f924da941f754c7a6d139bd662c2d0a3

Observation 24ae827a-277c-43c0-9b83-ff267a4b87f0 · outbound

This paper cites Social Attention for Autonomous Decision-Making in Dense Traffic.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Social Attention for Autonomous Decision-Making in Dense Traffic

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:11.091194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:11.091194Z digest=sha256:06f52a3f73163f40a4cc1e21b41ca5be8fa082bae75f1913fb3d6bd9f4d2d39d

Observation 634a9400-a604-4c21-ab8e-e6b1a4d0976c · outbound

This paper cites A multi-task reinforcement learning approach for navigating unsignalized intersections,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections A multi-task reinforcement learning approach for navigating unsignalized intersections,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:19.189396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:11.215996Z digest=sha256:9987029091bda810525beaad05f05ced335393d65d76c3ca64ab0777c69963f8

Observation cf90fdee-691c-4a1c-9e0f-19d1e152bec2 · outbound

This paper cites Pomdp and hierarchical options mdp with continuous actions for autonomous driving at intersections,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Pomdp and hierarchical options mdp with continuous actions for autonomous driving at intersections,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.944152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:11.385053Z digest=sha256:b3298285911232582d662d032e6b47b1805f0b86cde34cbca706dbdfb97c1757

Observation 0ba67e87-a373-4f1b-b026-cfd37ff24656 · outbound

This paper cites Behavior Planning at Urban Intersections through Hierarchical Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Behavior Planning at Urban Intersections through Hierarchical Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:15.806518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:11.506469Z digest=sha256:5e9e25d7c1c74a4a6ba8740e6aad87e9d5ca477fd883f31ecf33109c15bc3698

Observation bbac736c-efe4-496e-b0c8-7f9c88380e02 · outbound

This paper cites Action and trajectory planning for urban autonomous driving with hierarchical reinforcement learning,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Action and trajectory planning for urban autonomous driving with hierarchical reinforcement learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:11.624122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:11.624122Z digest=sha256:a462e737697b018444e7a33758698f739939ed37a715fe7255adf7afd4d0237a

Observation 6bac7e8b-6c20-4214-9cb5-17bd944f04bb · outbound

This paper cites Safe reinforcement learning on autonomous vehicles,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Safe reinforcement learning on autonomous vehicles,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.721818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:11.854782Z digest=sha256:b46c8501813a9004cc5e9897329f73e7abe5e819b86555df91a8a7289f8500a2

Observation eb576bdf-531c-4c0f-b9fd-d411cf778a2b · outbound

This paper cites Learning to Navigate Intersections with Unsupervised Driver Trait Inference.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning to Navigate Intersections with Unsupervised Driver Trait Inference

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:15.505005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:11.986404Z digest=sha256:3aba402e001797a35b777f4bd286b55c2dde0f76ed68d403a972cc8ca7a16006

Observation 6bc65589-4427-468b-a705-8815a3094f37 · outbound

This paper cites Human-like decision making at unsignalized intersections using social value ori- entation,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Human-like decision making at unsignalized intersections using social value ori- entation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.415770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:12.093783Z digest=sha256:beb568d6f0ae6c74f58b7f285efeab84ed1a6b385f1b795edd0602c947494ee4

Observation 91f99a7a-e874-4936-9487-37a3581c1db0 · outbound

This paper cites Interactive autonomous navigation with internal state inference and interactivity estimation,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Interactive autonomous navigation with internal state inference and interactivity estimation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.127625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:12.243166Z digest=sha256:b907df5b1b69d2f13cbf1539019b87e3610de5085fe7dd810b3a03745bd4503f

Observation d8417676-b812-4f2e-ab62-313a174f5141 · outbound

This paper cites Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.849826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:12.391343Z digest=sha256:74b358399e854f0598f717e726203181a3da59a892f37abedc1a6e6e5e2fec3f

Observation bcaaddc6-da14-4cf4-9159-1a266599e13a · outbound

This paper cites Feudal reinforcement learning,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Feudal reinforcement learning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.500305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.500305Z digest=sha256:f08aa0f497eb8d3d54177a9f41c558b2e021897203b3c6ac0e36ddad786f937d

Observation b1b5c3f8-d11b-48c9-84c0-67c1bff50f6f · outbound

This paper cites Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.644637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.644637Z digest=sha256:389b7bb3f9d132fd6996b2abbe28be74ec322578687dfcd61e10255783e9bc9f

Observation 1a17800a-5966-4d80-8d93-b589c280009d · outbound

This paper cites Data-Efficient Hierarchical Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Data-Efficient Hierarchical Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.786912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.786912Z digest=sha256:ed41173c16ab484dbf16979a9bcf9b605383813e6be8903a3e7d1f0c0a7fdaea

Observation 77b7febc-4a27-4da2-bba3-9c65c30fe74c · outbound

This paper cites Learning Multi-Level Hierarchies with Hindsight.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Multi-Level Hierarchies with Hindsight

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.947806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.947806Z digest=sha256:39dac277bb17f31a6a7ecf6817ad355c2f968e3fced991d94c99edb091acae5a

Observation a06babb2-978f-4762-b0cd-9c6a32bb8c83 · outbound

This paper cites Multipath: Multiple probabilistic anchor trajectory hypotheses for behavior prediction,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Multipath: Multiple probabilistic anchor trajectory hypotheses for behavior prediction,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.642287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:13.093012Z digest=sha256:a368862ae13ca14653890634e17d2d50f5efde0e3f8281ad416ab74234bd3fb1

Observation 6461e5fb-e054-4672-83cf-106621710998 · outbound

This paper cites Learning Lane Graph Representations for Motion Forecasting.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Lane Graph Representations for Motion Forecasting

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:15.123765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:13.400237Z digest=sha256:a08ed29c43863bad0f19d65095f6c871b1374ec0dd38965a11dcd7722dc4df66

Observation 42429821-ae29-4de1-b089-6cf453b50753 · outbound

This paper cites TNT: Target-driveN Trajectory Prediction.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections TNT: Target-driveN Trajectory Prediction

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.523673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.523673Z digest=sha256:9a9688c68dc9dc34b2e56ea8dc168ec0bf95e321b4cdfd75035e7189ceb98b43

Observation c66819fc-126f-4661-a848-982f4ad85d6f · outbound

This paper cites an unresolved cited work.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.652077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.652077Z digest=sha256:2f9493a6142ee09e7a49e0d41be43af6aa1a743cf0a8b247192857582aa538a3

Observation 45fc5354-722d-43df-b207-f9136d67e99a · outbound

This paper cites Learning Interaction-aware Motion Prediction Model for Decision-making in Autonomous Driving.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Interaction-aware Motion Prediction Model for Decision-making in Autonomous Driving

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.760119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.760119Z digest=sha256:9485d02de82c6792a8dd1a6834b260a6d95b131a7a7edb836d4c2f69f1404bf6

Observation a999beb6-da79-46da-aaed-6ccd82a66ce4 · outbound

This paper cites Scene Transformer: A unified architecture for predicting multiple agent trajectories.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Scene Transformer: A unified architecture for predicting multiple agent trajectories

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.842331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.842331Z digest=sha256:b50e2bf01519f431dc2bfb04bdbc995ac3ccb4a4c31700aa496e594486b5ef3d

Observation b785a379-6c96-4f74-bb76-b26ffe255d2d · outbound

This paper cites VectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representation.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections VectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.943431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.943431Z digest=sha256:4bcf221b4a12a3701611d9f54d0f24f9d4bfae2a2a739476de615b33b18541be

Observation e6714b33-faaf-4814-917b-6be7996d76dd · outbound

This paper cites Attention Is All You Need.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Attention Is All You Need

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.075473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.075473Z digest=sha256:140e68a11476523f5b626a18e4635bf5c3c8233c1f3ccde226932b7c266cbb5a

Observation dbe86f36-d6bf-4406-b8f7-077166cb340c · outbound

This paper cites Separating axis theorem for oriented bounding boxes,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Separating axis theorem for oriented bounding boxes,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.396641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:14.202926Z digest=sha256:0a8346744cf60587a8670e36b08c7361116d1f68d39eca438be91bf506d413e2

Observation 1a012cb7-39e4-4159-b9e1-ff95033a6cce · outbound

This paper cites Proximal Policy Optimization Algorithms.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.419292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.419292Z digest=sha256:3288d321866efb46a98ad81472a977856a3c2a895470c179ee4f75e71a8e4ac1

Observation 21ce2981-ed5c-403d-a2ca-899d11809ae8 · outbound

This paper cites SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.539948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.539948Z digest=sha256:5840be3def631330257c68181feb1776d831f42314ae111cc927f3e962bba2df

Observation 14840db5-cc5f-4874-bc3a-3300ae42800d · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Playing Atari with Deep Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.675888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.675888Z digest=sha256:1003a8bb6596bb0bf2ae2dd9d07893de8fd8d58ea4e69db50b9ffdcd60c3f80c

Observation 7b6c6c6b-ad19-4b2e-8a59-d6bf083e2595 · outbound

This paper cites Soft Actor-Critic for Discrete Action Settings.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Soft Actor-Critic for Discrete Action Settings

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.797853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.797853Z digest=sha256:e06cb044dc098ec8977ad4b44ff2ade2dc44b6e4136b661a627fd065bf82eeb2

Observation 8a9c4364-e1e3-4d73-9c04-347d07237822 · outbound

This paper cites Available: https://api.semanticscholar.org/CorpusID: 125587700.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Available: https://api.semanticscholar.org/CorpusID: 125587700

Reference 2009

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.124427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:49:14.338878Z digest=sha256:491d59bd9ebb3d7cea0806d39e9691ac7867f44e7a309798706507ccac2e4af0

Observation 4b2a6ee0-ded9-4b99-b15c-865d88d30efe · outbound

This paper cites End to End Learning for Self-Driving Cars.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections End to End Learning for Self-Driving Cars

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:10.555549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:10.555549Z digest=sha256:e364687cf02824e099a1668dee218a45ad3ab75feced5f9f549a810ec0b14274

Observation 09da6045-488f-4dcb-ab9f-b449363a2107 · outbound

This paper cites MultiPath: Multiple Probabilistic Anchor Trajectory Hypotheses for Behavior Prediction.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections MultiPath: Multiple Probabilistic Anchor Trajectory Hypotheses for Behavior Prediction

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.253143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.253143Z digest=sha256:9504cd1f2cd23a55a490038fd1684d75b604cc46e62fb7b99543b5596ab2538c

Observation e0ee5f6a-6b20-4728-b96b-b6eb5f78f3c2 · outbound

This paper cites Action and Trajectory Planning for Urban Autonomous Driving with Hierarchical Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Action and Trajectory Planning for Urban Autonomous Driving with Hierarchical Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:11.735202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:11.735202Z digest=sha256:809f631fe06b35629ee9b76db76e6715946d713be3e6b67cd04cc11ca8ffc68b

Pith citing papers

No inbound Pith citation observations are available.