Pith. sign in

Paper Citation Record · LEDGER

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.20671.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20671 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:42.129182Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact2
  • verified fuzzy14
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d817240d-38a3-4af2-b1f8-cc32045505ad · outbound

This paper cites Reincarnating reinforcement learning: Reusing prior computation to accelerate progress.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Reincarnating reinforcement learning: Reusing prior computation to accelerate progress

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.139343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.139343Z digest=sha256:2c441ec403b39315ffb4b8437aa1848e93f74e08e422708ea3b7a5b9f582ebff

Observation f8430a45-e6ef-49df-9ff3-396f11e51c05 · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.218896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.218896Z digest=sha256:51c5fda70a2d2762656be407d363007392f621352267fc7381e2347bb73a3101

Observation 22846068-a9f4-436b-98d3-e1a47af968fe · outbound

This paper cites OpenAI Gym.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.377259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.377259Z digest=sha256:66f48eafaef7954046178879e9cc66b3771041f49115806aff79908c9fb8a6ff

Observation 670978e9-6d84-46a8-99d0-96f2821570bf · outbound

This paper cites Exploration by Random Network Distillation.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Exploration by Random Network Distillation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.484831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.484831Z digest=sha256:71d921b7cc0ab38132e9cc9fe42a56a1b514d9f7c0bed71076da44cfa8880b70

Observation 62ebb324-b823-48ac-91bf-b85644f567ce · outbound

This paper cites Imitation learning from vague feedback.Advances in Neural Information Processing Systems, 36:48275–48292, 2023.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Imitation learning from vague feedback.Advances in Neural Information Processing Systems, 36:48275–48292, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:48.333620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:37.615663Z digest=sha256:88db7babbb2aadf6429096f1e695e5c072c4003c557ced70ae9dcdbfe7c96279

Observation e536404e-ff62-4905-96ff-13b9d3dd7b60 · outbound

This paper cites Statemask: Explaining deep reinforcement learning through state mask.Advances in Neural Information Processing Systems, 36:62457–62487, 2023.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Statemask: Explaining deep reinforcement learning through state mask.Advances in Neural Information Processing Systems, 36:62457–62487, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:48.142610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:37.771974Z digest=sha256:0cb348a37903db637879d06cab80b57aed6142f8f3112029dbddc19e1913a203

Observation 65fddf26-7569-45a7-9c3f-12e0a2611be7 · outbound

This paper cites RICE: Breaking Through the Training Bottlenecks of Reinforcement Learning with Explanation.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation RICE: Breaking Through the Training Bottlenecks of Reinforcement Learning with Explanation

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:53:43.193960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:37.920161Z digest=sha256:70ac9457672e896d27de70e546a39747c87943550a93870ddb680293c76307cb

Observation 3174fb9c-6810-4c86-b4d3-fe463fc0f735 · outbound

This paper cites Using Natural Language for Reward Shaping in Reinforcement Learning.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Using Natural Language for Reward Shaping in Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.045573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.045573Z digest=sha256:72aa16c96396cb5cfc07e2ccb1ee613ca43c445741e3aad20759a4cedbeb0cd5

Observation 1b974b84-8aec-4df4-ba8d-14f2524b65c8 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:47.935363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:38.171732Z digest=sha256:858918359df1fb571d75438837233680ec8be557b0e90444c6d2312061eafb1e

Observation e34f203c-d287-4e5f-80d6-ce50ded01619 · outbound

This paper cites Edge: Explaining deep reinforcement learning policies.Advances in Neural Information Processing Systems, 34:12222–12236, 2021.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Edge: Explaining deep reinforcement learning policies.Advances in Neural Information Processing Systems, 34:12222–12236, 2021

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.739766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:38.341754Z digest=sha256:35fdc568e734cd986de7c14e1f3ad7104c0e0e62b85c0658440944798d5f25d7

Observation ba18b94e-e4b3-42c1-982e-8ea44066974b · outbound

This paper cites Uncertainty-aware reinforcement learning for autonomous driving with multimodal digital driver guidance.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Uncertainty-aware reinforcement learning for autonomous driving with multimodal digital driver guidance

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.568713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:38.455639Z digest=sha256:bcfaa537eb0fe026d1277f93482580bb3b06d42b6663058c8a27d5dec53cac71

Observation 4812f92e-cc31-4232-831c-b9f281382fe4 · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.626574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.626574Z digest=sha256:b7542970be5d81d96c1caf8fb3efda2de0a47ed2ed2e388fc54fdd87f579737c

Observation e2c8dac2-a844-4e4f-a062-14f1f34ff23c · outbound

This paper cites Reward Design with Language Models.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Reward Design with Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.843038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.843038Z digest=sha256:b37c543e5c836453622efea0e127ab09e7194c63f05ec98b76b913c21dd3e327

Observation 693244ae-1063-432e-a529-52d44d492358 · outbound

This paper cites A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.995783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.995783Z digest=sha256:2d2cb0d08051064451fb60e1a1cc9c28c15bad5f61bc711b83f65f187b1ca34b

Observation 12681bee-52c3-4902-933a-a418ab9f68cb · outbound

This paper cites Traj-llm: A new exploration for empowering trajectory prediction with pre-trained large language models.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Traj-llm: A new exploration for empowering trajectory prediction with pre-trained large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.327917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:39.120232Z digest=sha256:24279c1b31e9829403051a525bbc6fa398a808001d1ba508c0e954a8357fa789

Observation 2713381e-02ec-4d0b-b4b3-831dc1b409e1 · outbound

This paper cites Pre-trained language models for interactive decision-making.Advances in Neural Information Processing Systems, 35:31199–31212, 2022.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Pre-trained language models for interactive decision-making.Advances in Neural Information Processing Systems, 35:31199–31212, 2022

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.249217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.249217Z digest=sha256:ecd273a90cda22519c61c4167f4cd5a9eed79c28cc1e1197cb73756fb83b6758

Observation ee228267-291e-4941-8d23-1deac2a35ed2 · outbound

This paper cites Utility: Utilizing explainable reinforcement learning to improve reinforcement learning.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Utility: Utilizing explainable reinforcement learning to improve reinforcement learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.105048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:39.382941Z digest=sha256:aa24aef7e5b59c8c170326f1f7a9026c3fa29a0db20a4e4365d6234cb1f88e53

Observation db71e279-304a-40eb-babd-6ef30e623a54 · outbound

This paper cites Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:53:42.823242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:39.462927Z digest=sha256:276919780d1bec2a922b277891ee24735326a4bbbbdf022f9f3beae3ab37f426

Observation a91e3f3f-b9af-4f4d-82a7-4c258e4dd159 · outbound

This paper cites Episodic Curiosity through Reachability.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Episodic Curiosity through Reachability

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.550890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.550890Z digest=sha256:3950841b8f501222ccddb39005f10f7997381cdcc5715067c922ee00876f42b8

Observation 01195cea-d6ab-40a3-a614-170d2368343d · outbound

This paper cites Proximal Policy Optimization Algorithms.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Proximal Policy Optimization Algorithms

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.735995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.735995Z digest=sha256:d163e9aa24608e1ff4c94245db0e6ed5d6cfa4ac317b4848280cd3c0cda17e1f

Observation 06068cad-f92a-42ea-8564-3df367bfa8d6 · outbound

This paper cites Perceiver-actor: A multi-task transformer for robotic manipulation.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Perceiver-actor: A multi-task transformer for robotic manipulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.913463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.913463Z digest=sha256:a33616ddf25f07782d1245eb26e55a26771550a97c4b6fba5cfc13d8e9eafce5

Observation 260cef10-6f61-4bbd-9f6a-b10b75bc6283 · outbound

This paper cites Joint rebalancing and charging for shared electric micromobility vehicles with energy-informed demand.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Joint rebalancing and charging for shared electric micromobility vehicles with energy-informed demand

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:46.732568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:40.037041Z digest=sha256:cafe82d08858edd47dfaf95b96eccbca59f93043e7a5ab379c2bbd010458f327

Observation 3c59896d-b416-4b67-bc3b-ffc9097a5b5e · outbound

This paper cites Mujoco: A physics engine for model-based control.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Mujoco: A physics engine for model-based control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.161934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.161934Z digest=sha256:95598bf097c2f2c2cf25fb13d0aba7f541ea978843704f283ba5f5be6d938a66

Observation 02e1134c-1b30-4393-9a81-ec0766101ccc · outbound

This paper cites Correct me if i’m wrong: Using non-experts to repair reinforcement learning policies.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Correct me if i’m wrong: Using non-experts to repair reinforcement learning policies

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:46.458670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:40.295047Z digest=sha256:56e5a800a4387ed017c6a6328a77adc022fc2e8f0263f2f0ef970c6e6f2302ad

Observation aeeaff42-2518-49ff-9fcc-b8685f8bdd45 · outbound

This paper cites Grandmaster level in starcraft ii using multi-agent reinforcement learning.nature, 575(7782):350–354, 2019.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Grandmaster level in starcraft ii using multi-agent reinforcement learning.nature, 575(7782):350–354, 2019

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.417064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.417064Z digest=sha256:2c7c82ddd1bbee9e3e5ddfe5891ab5eb0e1dc51583d33345eed7d046c3a64294

Observation 52c22c42-2152-43e2-87f6-f02544f078e2 · outbound

This paper cites STeCa: Step-level Trajectory Calibration for LLM Agent Learning.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation STeCa: Step-level Trajectory Calibration for LLM Agent Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.544944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.544944Z digest=sha256:b1fa9ae5c74e3e9ae47a5e199b063a0a7b0866652096dfc24370a6f359d9e095

Observation afa650bf-3544-4a5a-b0ec-b02436189bff · outbound

This paper cites Read and reap the rewards: Learning to play atari with the help of instruction manuals.Advances in Neural Information Processing Systems, 36:1009–1023, 2023.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Read and reap the rewards: Learning to play atari with the help of instruction manuals.Advances in Neural Information Processing Systems, 36:1009–1023, 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:46.197516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:40.618574Z digest=sha256:bf76914ce432e2b5c8208a423ee994f22f61b7512c4ac0722656315b26043716

Observation 99f315af-1619-4fde-a573-5a9787d0ae87 · outbound

This paper cites Keep CALM and Explore: Language Models for Action Generation in Text-based Games.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Keep CALM and Explore: Language Models for Action Generation in Text-based Games

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.731232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.731232Z digest=sha256:f19b14720a9fc76318d109918ad2de09970f4b5162339f404a8093d5c242c6c8

Observation 100c3049-4c9c-406e-9c5d-55f8c72b669d · outbound

This paper cites Offline Imitation Learning Through Graph Search and Retrieval.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Offline Imitation Learning Through Graph Search and Retrieval

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.827370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.827370Z digest=sha256:f9835f85246f90f210f5f9b9dadd679cdf15867874f646aa7598221454b268f5

Observation 846a6096-7ead-4efe-b425-75e07df69db1 · outbound

This paper cites {AIRS}: Expla- nation for deep reinforcement learning based security applications.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation {AIRS}: Expla- nation for deep reinforcement learning based security applications

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:45.933464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:40.914759Z digest=sha256:ea05799773905a5fe83b3318216388e022e1b8d3afdcc32a3ab0f4a36131e4ac

Observation af353fed-4776-48f1-b44b-244220b05b17 · outbound

This paper cites Language to Rewards for Robotic Skill Synthesis.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Language to Rewards for Robotic Skill Synthesis

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:41.067156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:41.067156Z digest=sha256:73a650c5847117b7b3845be1b0a3b2ca1bad1c71b8939f28897c43677d381024

Observation 06958167-1ff9-409e-a8d3-763007f96967 · outbound

This paper cites + str(e) +.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation + str(e) +

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:45.573248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.183085Z digest=sha256:7082b00075cf8d29bda71360362432525606834fe6e23faa93402cfe1f951664

Observation 7534699e-44c3-4777-affd-0503001fe5ef · outbound

This paper cites We should introduce a mechanism to reinforce learning during critical times while undermining actions leading to unfavorable outcomes.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation We should introduce a mechanism to reinforce learning during critical times while undermining actions leading to unfavorable outcomes

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:45.259951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.329697Z digest=sha256:f50e951f3ab79ec0a413b4219187c584dc5a7be8cecbe2559e51b89cc1c87174

Observation 3534af9b-9eb9-495c-aef9-08c1edb9c65d · outbound

This paper cites The policy should focus on maintaining positional advantage to intercept the ball efficiently.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation The policy should focus on maintaining positional advantage to intercept the ball efficiently

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:44.902909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.439086Z digest=sha256:2c3dfcea363084991a6ade4a6888c15560038b30fbba567c6dc4a9f2cf18c5b1

Observation 6cc0f9ff-df86-4f86-b4db-b42a43767c48 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:44.549204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.611895Z digest=sha256:493b991fea0bbc432b4b4dac9d2a893a925ab35135823a1d04130fd1f348464d

Observation 93396c86-d22d-4fff-aa84-14918688c819 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:44.293375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.728655Z digest=sha256:7abbb23ea3b67cbfb5b559ceaf3c60f54ce99d7dfd5ab092c7ada06f61abda9c

Observation 586c101f-69b3-4154-81a5-a5d5b6186390 · outbound

This paper cites Implementing a mechanism that prioritizes actions based on proximity to the ball in the x-coordinates can significantly improve action selection during critical moments.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Implementing a mechanism that prioritizes actions based on proximity to the ball in the x-coordinates can significantly improve action selection during critical moments

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:43.976039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.854397Z digest=sha256:a43a518fd4f40dc10a6b09011a59b16586110307d98244293a2076a62f861ee8

Observation d72a2abc-3c3f-4519-8eb0-cd57a5fa90e9 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:43.797107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:41.987874Z digest=sha256:f0d0b43c0a9e098ccafb9b723b15fbe163816199779f17f5150133f9478b3992

Observation 38ef5141-ee5b-463d-9b99-8f1a0e6391c9 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:43.615588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T13:53:42.129182Z digest=sha256:777f5007859b12f764354335f6d570092ece1376d5a99ca563a2d0a82abfdce8

Pith citing papers

No inbound Pith citation observations are available.