Pith. sign in

Paper Citation Record · LEDGER

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models

As of 23 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2501.05057.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05057 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:23:20.545675Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:13:41.473724Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:13:45.789156Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact1
  • verified fuzzy25
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9c0ec740-bba7-4405-8a47-d038f30db641 · outbound

This paper cites A Survey of Large Language Models.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A Survey of Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.319308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.319308Z digest=sha256:23000761f86ddbfe139bbb95863416d4208921c490a770b18b2446649e61be90

Observation 0f65d00d-0e1e-4896-8256-b0fe8a12b4c6 · outbound

This paper cites A survey on evaluation of large language models,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A survey on evaluation of large language models,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.325215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.325215Z digest=sha256:8ff28680dbb9ba8871df07f8a52de57bbcdde2698fac800d4e80b96a4b47491d

Observation ae648a41-b746-4658-b9ed-2f1b030fcbcf · outbound

This paper cites A survey on multimodal large language models for autonomous driving,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A survey on multimodal large language models for autonomous driving,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.555434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.330027Z digest=sha256:2dc3c42068075f4b6f48efa259243f32c93bdaa088058cb0f5a69da731556d8f

Observation a7e49d22-0725-4ac4-8da7-e1255f1f20db · outbound

This paper cites Deep reinforcement learning for autonomous driving: A survey,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Deep reinforcement learning for autonomous driving: A survey,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.334794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.334794Z digest=sha256:273b2cce3f56936a9b8e6fe6065a79c67cb8c57545f76f3271ebd74acf43568b

Observation ea9d91e9-c1f6-44ad-b7ef-964d33b18b0f · outbound

This paper cites Deep learning for safe autonomous driving: Current chal- lenges and future directions,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Deep learning for safe autonomous driving: Current chal- lenges and future directions,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.511219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.339503Z digest=sha256:3ed395e5bbb11eb4ff815d75d843e431bd6cd01c4a76f251d36442b2a76e4b5e

Observation 00b6254c-2a1f-4dad-bdc7-c479e461d9ca · outbound

This paper cites Deep learning-based vehicle behavior prediction for autonomous driving applications: A review,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Deep learning-based vehicle behavior prediction for autonomous driving applications: A review,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.343886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.343886Z digest=sha256:d1c2c8671bf928c93d64eca371feebe35703d2ce959299884585ee913e9df55b

Observation c5ea2442-f551-4468-b264-1789558119d7 · outbound

This paper cites Large scale interactive motion forecasting for autonomous driving: The waymo open motion dataset,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Large scale interactive motion forecasting for autonomous driving: The waymo open motion dataset,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.349589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.349589Z digest=sha256:95c32b66c8d11b4da0f8cc741ee7b66dd386000589ba518dc4b1f00348e7a13b

Observation 8e3956e3-dd62-4281-bbfe-c1c5f53e2d80 · outbound

This paper cites Behavior planning at urban in- tersections through hierarchical reinforcement learning,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Behavior planning at urban in- tersections through hierarchical reinforcement learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.460507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.353750Z digest=sha256:4ab85dc0e7662382e3c1cf067e8980b565f8e0626434c6e5bdc39f4ae4660510

Observation 15a9f599-75a0-4cc7-b189-f148df4aa625 · outbound

This paper cites Chance-aware lane change with high-level model predictive control through curriculum reinforcement learning,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Chance-aware lane change with high-level model predictive control through curriculum reinforcement learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.443324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.358319Z digest=sha256:0704677e887222c7c9b1ec91b8a04a0be3d2e771766ab98dd6b6c66e17418d62

Observation 0971f943-c18e-4f8d-9024-6027493be1e2 · outbound

This paper cites Reward-driven automated curriculum learning for interaction-aware self-driving at unsignalized intersections,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Reward-driven automated curriculum learning for interaction-aware self-driving at unsignalized intersections,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.427454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.362765Z digest=sha256:ffdb86afe2800d3e32b889f3c1a02b2c58c6c2cb2b8b8c3ce34d8bb6a2b59b05

Observation e9444779-07ff-435f-a925-3f364801783c · outbound

This paper cites The perils of trial-and-error reward design: misdesign through overfitting and invalid task specifications,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models The perils of trial-and-error reward design: misdesign through overfitting and invalid task specifications,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.411036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.366922Z digest=sha256:f39fb573bae122eff35a19ccd57e96c0a032f53f9543b5e2e0814a2d0573ba94

Observation 5da25052-e7f1-4be3-a9ea-93126015eb4e · outbound

This paper cites Learning to utilize shaping rewards: A new approach of reward shaping,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Learning to utilize shaping rewards: A new approach of reward shaping,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.371475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.371475Z digest=sha256:9aae859a934aa1504620ac16662e5d27ed282ca8446916f0387d5b8ea9c67005

Observation d40221e8-5493-491d-98e3-ca6f40d631bb · outbound

This paper cites an unresolved cited work.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:23:21.385705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.375702Z digest=sha256:742576f167b18a3d85687aada4709acc0ca731efdb859e6634e0429be75657b0

Observation cac0da9a-ac47-4e7d-8d83-803a0338381f · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.380158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.380158Z digest=sha256:763f8fb270c62b1019dbe54771fca93466201765e5a9a8c88ec6d267e21fd2f5

Observation 3d2f8eab-d19b-45b9-9233-9db41527a231 · outbound

This paper cites A review of reward functions for reinforcement learning in the context of autonomous driving,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A review of reward functions for reinforcement learning in the context of autonomous driving,

Reference 15

Resolution
verified exact
raw_fallback, observed 2026-08-10T21:23:20.946487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.385243Z digest=sha256:bc5475a682b5e33ac7e3e56af28e180736b850aca0e499ea9b9f57a9ac09914c

Observation c157aaa0-3b1c-412d-8874-db0dea0c39b8 · outbound

This paper cites Curriculum learning for reinforcement learning domains: A framework and survey,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum learning for reinforcement learning domains: A framework and survey,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.389739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.389739Z digest=sha256:050b9d54322a8ec90cab3a523dd756886ad4ccebbeb37c3b7b9029e38b6dde8a

Observation 6cf542e1-e1ce-49ae-8b95-36c87579ab12 · outbound

This paper cites Curriculum learning,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum learning,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.394404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.394404Z digest=sha256:040d822a75e63f9c7c91947b5a6f538a0326dd4f226d0cb55f1d1bb5fd905089

Observation e0b4a500-4286-494b-af53-d1f9019821ea · outbound

This paper cites Curriculum learning: A survey,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum learning: A survey,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.399013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.399013Z digest=sha256:06a203c9c1bcd53aaff4688b10dd5f3f1b8e3925a79f1327aa0320fd9223ec8f

Observation ffdddfd3-5628-497a-9d02-15dfb15d5822 · outbound

This paper cites Automated curriculum learning for neural networks,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Automated curriculum learning for neural networks,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.337953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.403685Z digest=sha256:20c1280fd6e985136c78c88653c0005633a6e210bb39f46f20e1e3d37729e991

Observation f02328a6-984c-4730-a035-0429e9ec7ae9 · outbound

This paper cites Robot parkour learning,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Robot parkour learning,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.322480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.408568Z digest=sha256:c12aa9a04b52751280adad066492478c5c75e0afa2b7eb4ce33ec327f82eb435

Observation 370b7dc6-e493-4a48-8431-264a9d48ecc8 · outbound

This paper cites Self-learned autonomous driving at unsignalized intersec- tions: A hierarchical reinforced learning approach for feasible decision- making,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Self-learned autonomous driving at unsignalized intersec- tions: A hierarchical reinforced learning approach for feasible decision- making,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.306218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.413416Z digest=sha256:83890a925b7c21c6fb574505ecdff8f3f5f153d949a205a1a3307f1943e76ffb

Observation a7db27e1-0d20-4fe9-8814-b205fbb474bd · outbound

This paper cites A multi-task reinforcement learning approach for navigating unsignalized intersec- tions,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A multi-task reinforcement learning approach for navigating unsignalized intersec- tions,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.289685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.418244Z digest=sha256:9bf9fe6e56066220aa5469135311e67aff88af66de24e6c628fbe885b1633788

Observation f1524a2a-24fb-40c8-9c58-4ca94b1c053e · outbound

This paper cites Multi-task reinforcement learning with context-based representations,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Multi-task reinforcement learning with context-based representations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.272809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.422944Z digest=sha256:668280a9a9b36b3278110dfff582881481edaa9d8ad63df9e57208f3181ff895

Observation 33c73be2-f491-47d7-9848-65a69a03f3c0 · outbound

This paper cites Multi-task safe reinforcement learning for navigating intersections in dense traffic,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Multi-task safe reinforcement learning for navigating intersections in dense traffic,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.427509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.427509Z digest=sha256:4dbbf91df47267fe693bdd852f3b78866a9f796deaa11573f853b5f94c0ffaa9

Observation b6b6c8c8-a115-4361-88db-723a0131e68a · outbound

This paper cites Learning robust rewards with adver- sarial inverse reinforcement learning,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Learning robust rewards with adver- sarial inverse reinforcement learning,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.246587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.431976Z digest=sha256:b3957f30020e980e6c8faa3d091bbe06f3c6e864c3d5ff2ae3f2a17a5cf5d08e

Observation 789bb713-f49d-43d3-b83a-ad2bf136fa4e · outbound

This paper cites A survey of inverse reinforcement learning: Challenges, methods and progress,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models A survey of inverse reinforcement learning: Challenges, methods and progress,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.231100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.436344Z digest=sha256:b0f2e6c31a00be80db038ab543968022f8836c65f7b67faac34371c438bd0119

Observation 28f613bb-6432-4abc-a2c8-3a3a042ffbf2 · outbound

This paper cites Driving in Real Life with Inverse Reinforcement Learning.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Driving in Real Life with Inverse Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.440936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.440936Z digest=sha256:7cf00952892051896e5bc0320a6601757cb7d510389bcd30622143d49215ce95

Observation b41c043b-1c5e-4adc-a821-4921d35ed981 · outbound

This paper cites Interaction- aware planning with deep inverse reinforcement learning for human- like autonomous driving in merge scenarios,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Interaction- aware planning with deep inverse reinforcement learning for human- like autonomous driving in merge scenarios,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.215211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.445727Z digest=sha256:7dfddbfb85a87951bfce396b444075f6dd3d418239cc92c05076a81e61922626

Observation 789b3bd2-a072-4580-9df5-9c3794684af9 · outbound

This paper cites Exploration-guided reward shaping for reinforcement learning under sparse rewards,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Exploration-guided reward shaping for reinforcement learning under sparse rewards,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.450615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.450615Z digest=sha256:d506e3bc2464b7d51d36bf8640ed3491fe57c23b6bdb1fcc096f14253bd9d028

Observation 5f299bd2-c4e6-4bba-8554-bc866c1063ab · outbound

This paper cites Reward design with language models,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Reward design with language models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.187953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.455169Z digest=sha256:4fb9b569f88d12a82697bf79e2900f0ade069979e7f806acd7a8552b0209cabc

Observation 446c984c-96db-47ba-b3e0-1cd406111cdf · outbound

This paper cites Auto mc-reward: Automated dense reward design with large language models for minecraft,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Auto mc-reward: Automated dense reward design with large language models for minecraft,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.171504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.459679Z digest=sha256:3daddc3ff7f75ec88bf43dd40e4ee8a339869ae05b2aaaa41df6e071bc1f8a54

Observation 59b10c27-b8e1-485b-910e-f3987c1abe18 · outbound

This paper cites Centralized cooperation for connected and automated vehicles at intersections by proximal policy optimization,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Centralized cooperation for connected and automated vehicles at intersections by proximal policy optimization,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.464397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.464397Z digest=sha256:806f988fca30502fafb6e1eefd48e3272a6bc6e115fbdaffdf6204671ade95ac

Observation c716c27f-ea00-4d90-869d-73f577d50742 · outbound

This paper cites Au- tonomous overtaking in Gran Turismo sport using curriculum rein- forcement learning,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Au- tonomous overtaking in Gran Turismo sport using curriculum rein- forcement learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.143917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.469056Z digest=sha256:c55720b34ce62cbe7284e504b5f5fa64f9da828be8d6f4bf04e39c21133dfb06

Observation 9cf8110e-3dcd-4e19-946f-b4a48bf8ecf8 · outbound

This paper cites Curriculum proximal policy optimization with stage-decaying clipping for self- driving at unsignalized intersections,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Curriculum proximal policy optimization with stage-decaying clipping for self- driving at unsignalized intersections,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.126296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.473999Z digest=sha256:165d2a7370d18d97d2752f80888281d9db8d6464e3f33eba039e8a952aee86cb

Observation e1aee4d8-f551-4c9a-8e0f-c30ce08419c8 · outbound

This paper cites Automatically generated curriculum based reinforcement learning for autonomous vehicles in urban environment,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Automatically generated curriculum based reinforcement learning for autonomous vehicles in urban environment,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.110126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.478595Z digest=sha256:b18094db4376c6a81d25448e695d346798899d965f8b86978655e54c53375302

Observation 0416e0e1-a419-4132-bc30-1490c7312021 · outbound

This paper cites State dropout-based curriculum reinforce- ment learning for self-driving at unsignalized intersections,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models State dropout-based curriculum reinforce- ment learning for self-driving at unsignalized intersections,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.093680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.483273Z digest=sha256:08f81a174b91ed45b6595dec3a38779b49d260061ece042eabe8191f1ddeea40

Observation cdb15ce1-6d61-499a-8a4d-533f928856c0 · outbound

This paper cites Dilu: A knowledge-driven approach to autonomous driving with large language models,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Dilu: A knowledge-driven approach to autonomous driving with large language models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.077356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.489082Z digest=sha256:fcae4a8ccc0db144049d4c6ba1b86e1e4799ae99e5267f336f46a3a5e292ba24

Observation e259233e-f6b9-4eec-900a-137ebd7b56ee · outbound

This paper cites Drivegpt4: Interpretable end-to-end autonomous driving via large language model,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Drivegpt4: Interpretable end-to-end autonomous driving via large language model,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.494368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.494368Z digest=sha256:fdeae29ebff58f22b246c125342704c735dedada48b2b6bbe686ba4cc4b50b9f

Observation 96792b4e-62bb-4162-aef5-c785c1a88822 · outbound

This paper cites Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.499346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.499346Z digest=sha256:8240de68010a24af73d1801491091758e8a5ee71dc47652b6b92c7bd7b55cd29

Observation e2754f77-3c41-4a8c-bb7b-857ce9da54cd · outbound

This paper cites DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.504100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.504100Z digest=sha256:8f7a32aaf54bdb0794a5521320e499fa7356266a9728e6427901ced76a8e659a

Observation 02c69e66-2a6d-44f3-a89e-e2256b3ab263 · outbound

This paper cites Language to 14 rewards for robotic skill synthesis,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Language to 14 rewards for robotic skill synthesis,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.052323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.508608Z digest=sha256:61ceaf10c65e4eb60326c09b5ad8912ec42fb5725ede6ad71271f547b0a56320

Observation 68c4d726-39aa-4b74-bb02-d46ad8466384 · outbound

This paper cites REvolve: Reward Evolution with Large Language Models using Human Feedback.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models REvolve: Reward Evolution with Large Language Models using Human Feedback

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.512620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.512620Z digest=sha256:9f4f32c7275eb28dce90e996c956032c89dd39943ce9c663625325bbfa1f7202

Observation ddcb794a-3c12-4229-af19-995f8f9f066a · outbound

This paper cites Human-centric Reward Optimization for Reinforcement Learning-based Automated Driving using Large Language Models.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Human-centric Reward Optimization for Reinforcement Learning-based Automated Driving using Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.517160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.517160Z digest=sha256:0dfd45fe9767420b5c70bfbabc881e3503d8c94676c559c4a350a439b80304a6

Observation ffc041d8-05e7-4446-bb53-c5b3fd2806a1 · outbound

This paper cites Dreureka: Language model guided sim-to-real transfer,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Dreureka: Language model guided sim-to-real transfer,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.037687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.521808Z digest=sha256:bf2f69a55c2f69b5ecac3143783b1acd518a3b41af31fb052b8ee2e1227f099c

Observation 5581ed92-4026-4e92-9517-87ced3da2592 · outbound

This paper cites CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models CurricuLLM: Automatic Task Curricula Design for Learning Complex Robot Skills using Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.526441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.526441Z digest=sha256:f88653de1ac2eb8db43f425f0983fe9a247d4541a6fcc4689fc8a2df7cb6b49a

Observation 578f496c-fa07-438b-99d2-7b878d2c17eb · outbound

This paper cites Autoreward: Closed-loop reward design with large language models for autonomous driving,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Autoreward: Closed-loop reward design with large language models for autonomous driving,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.022341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.531485Z digest=sha256:4c2c5e3746b3bdca23d09cebd69e78b6f7b0ec941ffbc7bb28b02bd2c49d720a

Observation 5daacbbd-efdb-4c6d-8b54-aad17f5a9e0d · outbound

This paper cites CARLA: An open urban driving simulator,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models CARLA: An open urban driving simulator,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:23:21.006570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-10T21:23:20.536281Z digest=sha256:509f29b57ab0def01480737cd3437fcc6d1bb4aea09403967ab0b665adbb64e0

Observation b92082a0-f68c-421f-b25e-6141fb0a4740 · outbound

This paper cites Proximal Policy Optimization Algorithms.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Proximal Policy Optimization Algorithms

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.540793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.540793Z digest=sha256:8a9feb384cef38c19feb482a453d57dcd84a0977c97a21cea38ac3023440a2a3

Observation 050329e2-104a-4e2a-a88b-f2352da2604c · outbound

This paper cites Casadi: a software framework for nonlinear optimization and optimal control,.

LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models Casadi: a software framework for nonlinear optimization and optimal control,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:23:20.545675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:23:20.545675Z digest=sha256:fad4f97b7273363602066d8e966028d91eecb7f91c401401a7e6fbdb8a5aab35

Pith citing papers

Observation eedf6962-2325-4712-84a7-9fae99fe9538 · inbound

HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving cites this paper.

HCRMP: A LLM-Hinted Contextual Reinforcement Learning Framework for Autonomous Driving LearningFlow: Automated Policy Learning Workflow for Urban Driving with Large Language Models

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:13:45.844889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-07T15:13:41.473724Z digest=sha256:65659443d094b4352b0bd8a9f8569a2e2e1196cc1e2b9f2517e257a5c9434e3b