Pith. sign in

Paper Citation Record · LEDGER

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.20671.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20671 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:53:42.129182Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact2
  • verified fuzzy14
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d817240d-38a3-4af2-b1f8-cc32045505ad · outbound

This paper cites Reincarnating reinforcement learning: Reusing prior computation to accelerate progress.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Reincarnating reinforcement learning: Reusing prior computation to accelerate progress

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.139343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.139343Z digest=sha256:2c441ec403b39315ffb4b8437aa1848e93f74e08e422708ea3b7a5b9f582ebff

Observation f8430a45-e6ef-49df-9ff3-396f11e51c05 · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.218896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.218896Z digest=sha256:51c5fda70a2d2762656be407d363007392f621352267fc7381e2347bb73a3101

Observation 22846068-a9f4-436b-98d3-e1a47af968fe · outbound

This paper cites OpenAI Gym.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.377259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.377259Z digest=sha256:66f48eafaef7954046178879e9cc66b3771041f49115806aff79908c9fb8a6ff

Observation 670978e9-6d84-46a8-99d0-96f2821570bf · outbound

This paper cites Exploration by Random Network Distillation.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Exploration by Random Network Distillation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:37.484831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:37.484831Z digest=sha256:71d921b7cc0ab38132e9cc9fe42a56a1b514d9f7c0bed71076da44cfa8880b70

Observation 62ebb324-b823-48ac-91bf-b85644f567ce · outbound

This paper cites Imitation learning from vague feedback.Advances in Neural Information Processing Systems, 36:48275–48292, 2023.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Imitation learning from vague feedback.Advances in Neural Information Processing Systems, 36:48275–48292, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:48.333620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:37.615663Z digest=sha256:e4b15ac6bc1ecbbae441c0404ce6c428c31e066f36047c7540758f722818ab6c

Observation e536404e-ff62-4905-96ff-13b9d3dd7b60 · outbound

This paper cites Statemask: Explaining deep reinforcement learning through state mask.Advances in Neural Information Processing Systems, 36:62457–62487, 2023.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Statemask: Explaining deep reinforcement learning through state mask.Advances in Neural Information Processing Systems, 36:62457–62487, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:48.142610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:37.771974Z digest=sha256:c39de4889cb175266fec14a44cdfb37828d4a953b83f7156000f5cebc6d373e9

Observation 65fddf26-7569-45a7-9c3f-12e0a2611be7 · outbound

This paper cites RICE: Breaking Through the Training Bottlenecks of Reinforcement Learning with Explanation.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation RICE: Breaking Through the Training Bottlenecks of Reinforcement Learning with Explanation

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:53:43.193960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:37.920161Z digest=sha256:e1932c3447399acd821370be7307f2737bca0b60ada9a348d1b4d2a816cfd88b

Observation 3174fb9c-6810-4c86-b4d3-fe463fc0f735 · outbound

This paper cites Using Natural Language for Reward Shaping in Reinforcement Learning.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Using Natural Language for Reward Shaping in Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.045573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.045573Z digest=sha256:72aa16c96396cb5cfc07e2ccb1ee613ca43c445741e3aad20759a4cedbeb0cd5

Observation 1b974b84-8aec-4df4-ba8d-14f2524b65c8 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:47.935363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:38.171732Z digest=sha256:c66c308c795e4755410381720f9307df101e998b23f39cad2b6d0685f3b7733f

Observation e34f203c-d287-4e5f-80d6-ce50ded01619 · outbound

This paper cites Edge: Explaining deep reinforcement learning policies.Advances in Neural Information Processing Systems, 34:12222–12236, 2021.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Edge: Explaining deep reinforcement learning policies.Advances in Neural Information Processing Systems, 34:12222–12236, 2021

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.739766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:38.341754Z digest=sha256:59bee684bd367c0da63f1e9dec738550f943bc2814424b76910700cfd184bf09

Observation ba18b94e-e4b3-42c1-982e-8ea44066974b · outbound

This paper cites Uncertainty-aware reinforcement learning for autonomous driving with multimodal digital driver guidance.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Uncertainty-aware reinforcement learning for autonomous driving with multimodal digital driver guidance

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.568713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:38.455639Z digest=sha256:5cfc8e6a02c225b9034ef6a305b03ff8526621db8796aa38ab0a3b5959c4af7f

Observation 4812f92e-cc31-4232-831c-b9f281382fe4 · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.626574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.626574Z digest=sha256:43ad5a17a0eaec34c5e96bf34643acdad5bf3c02de11bfa7758c76dbd5cd1e8c

Observation e2c8dac2-a844-4e4f-a062-14f1f34ff23c · outbound

This paper cites Reward Design with Language Models.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Reward Design with Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.843038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.843038Z digest=sha256:b37c543e5c836453622efea0e127ab09e7194c63f05ec98b76b913c21dd3e327

Observation 693244ae-1063-432e-a529-52d44d492358 · outbound

This paper cites A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation A Survey On Enhancing Reinforcement Learning in Complex Environments: Insights from Human and LLM Feedback

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:38.995783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:38.995783Z digest=sha256:2d2cb0d08051064451fb60e1a1cc9c28c15bad5f61bc711b83f65f187b1ca34b

Observation 12681bee-52c3-4902-933a-a418ab9f68cb · outbound

This paper cites Traj-llm: A new exploration for empowering trajectory prediction with pre-trained large language models.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Traj-llm: A new exploration for empowering trajectory prediction with pre-trained large language models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.327917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:39.120232Z digest=sha256:44bbd0ccc34cbea39e0c296b8d1abe58102d3105b27478754858e6a6e1407548

Observation 2713381e-02ec-4d0b-b4b3-831dc1b409e1 · outbound

This paper cites Pre-trained language models for interactive decision-making.Advances in Neural Information Processing Systems, 35:31199–31212, 2022.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Pre-trained language models for interactive decision-making.Advances in Neural Information Processing Systems, 35:31199–31212, 2022

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.249217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.249217Z digest=sha256:ecd273a90cda22519c61c4167f4cd5a9eed79c28cc1e1197cb73756fb83b6758

Observation ee228267-291e-4941-8d23-1deac2a35ed2 · outbound

This paper cites Utility: Utilizing explainable reinforcement learning to improve reinforcement learning.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Utility: Utilizing explainable reinforcement learning to improve reinforcement learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:47.105048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:39.382941Z digest=sha256:2680ccdf53b4b707e04dc013bffe80ed07ba1f7cb867d7f72a9b161222a9b9f8

Observation db71e279-304a-40eb-babd-6ef30e623a54 · outbound

This paper cites Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Mapping out the Space of Human Feedback for Reinforcement Learning: A Conceptual Framework

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:53:42.823242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:39.462927Z digest=sha256:043fcd8d2836f2f271d20d8b7f5ee6130e09889999f2b70850a045e379b93c09

Observation a91e3f3f-b9af-4f4d-82a7-4c258e4dd159 · outbound

This paper cites Episodic Curiosity through Reachability.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Episodic Curiosity through Reachability

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.550890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.550890Z digest=sha256:3950841b8f501222ccddb39005f10f7997381cdcc5715067c922ee00876f42b8

Observation 01195cea-d6ab-40a3-a614-170d2368343d · outbound

This paper cites Proximal Policy Optimization Algorithms.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Proximal Policy Optimization Algorithms

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.735995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.735995Z digest=sha256:d163e9aa24608e1ff4c94245db0e6ed5d6cfa4ac317b4848280cd3c0cda17e1f

Observation 06068cad-f92a-42ea-8564-3df367bfa8d6 · outbound

This paper cites Perceiver-actor: A multi-task transformer for robotic manipulation.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Perceiver-actor: A multi-task transformer for robotic manipulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:39.913463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:39.913463Z digest=sha256:a33616ddf25f07782d1245eb26e55a26771550a97c4b6fba5cfc13d8e9eafce5

Observation 260cef10-6f61-4bbd-9f6a-b10b75bc6283 · outbound

This paper cites Joint rebalancing and charging for shared electric micromobility vehicles with energy-informed demand.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Joint rebalancing and charging for shared electric micromobility vehicles with energy-informed demand

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:46.732568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:40.037041Z digest=sha256:7bea011da3a0e0b31134938727c5087ba3e8096f1ec33b084b71958cbcffc073

Observation 3c59896d-b416-4b67-bc3b-ffc9097a5b5e · outbound

This paper cites Mujoco: A physics engine for model-based control.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Mujoco: A physics engine for model-based control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.161934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.161934Z digest=sha256:95598bf097c2f2c2cf25fb13d0aba7f541ea978843704f283ba5f5be6d938a66

Observation 02e1134c-1b30-4393-9a81-ec0766101ccc · outbound

This paper cites Correct me if i’m wrong: Using non-experts to repair reinforcement learning policies.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Correct me if i’m wrong: Using non-experts to repair reinforcement learning policies

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:46.458670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:40.295047Z digest=sha256:b1754118970ae5cb9d742d369a16b31f4febd27134ca32a6de53acab161a4967

Observation aeeaff42-2518-49ff-9fcc-b8685f8bdd45 · outbound

This paper cites Grandmaster level in starcraft ii using multi-agent reinforcement learning.nature, 575(7782):350–354, 2019.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Grandmaster level in starcraft ii using multi-agent reinforcement learning.nature, 575(7782):350–354, 2019

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.417064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.417064Z digest=sha256:2c7c82ddd1bbee9e3e5ddfe5891ab5eb0e1dc51583d33345eed7d046c3a64294

Observation 52c22c42-2152-43e2-87f6-f02544f078e2 · outbound

This paper cites STeCa: Step-level Trajectory Calibration for LLM Agent Learning.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation STeCa: Step-level Trajectory Calibration for LLM Agent Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.544944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.544944Z digest=sha256:b1fa9ae5c74e3e9ae47a5e199b063a0a7b0866652096dfc24370a6f359d9e095

Observation afa650bf-3544-4a5a-b0ec-b02436189bff · outbound

This paper cites Read and reap the rewards: Learning to play atari with the help of instruction manuals.Advances in Neural Information Processing Systems, 36:1009–1023, 2023.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Read and reap the rewards: Learning to play atari with the help of instruction manuals.Advances in Neural Information Processing Systems, 36:1009–1023, 2023

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:46.197516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:40.618574Z digest=sha256:32ef16fd01f292e28b0fb806dc808711089fe4a07e3f340bc4771e445cf00b1d

Observation 99f315af-1619-4fde-a573-5a9787d0ae87 · outbound

This paper cites Keep CALM and Explore: Language Models for Action Generation in Text-based Games.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Keep CALM and Explore: Language Models for Action Generation in Text-based Games

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.731232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.731232Z digest=sha256:f19b14720a9fc76318d109918ad2de09970f4b5162339f404a8093d5c242c6c8

Observation 100c3049-4c9c-406e-9c5d-55f8c72b669d · outbound

This paper cites Offline Imitation Learning Through Graph Search and Retrieval.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Offline Imitation Learning Through Graph Search and Retrieval

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:40.827370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:40.827370Z digest=sha256:f9835f85246f90f210f5f9b9dadd679cdf15867874f646aa7598221454b268f5

Observation 846a6096-7ead-4efe-b425-75e07df69db1 · outbound

This paper cites {AIRS}: Expla- nation for deep reinforcement learning based security applications.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation {AIRS}: Expla- nation for deep reinforcement learning based security applications

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:45.933464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:40.914759Z digest=sha256:dfa3ef86a51f90ade53d3f1752d0b4cb79132a1a4115441adca7e3427c1be5f1

Observation af353fed-4776-48f1-b44b-244220b05b17 · outbound

This paper cites Language to Rewards for Robotic Skill Synthesis.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Language to Rewards for Robotic Skill Synthesis

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:41.067156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:41.067156Z digest=sha256:73a650c5847117b7b3845be1b0a3b2ca1bad1c71b8939f28897c43677d381024

Observation 06958167-1ff9-409e-a8d3-763007f96967 · outbound

This paper cites + str(e) +.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation + str(e) +

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:45.573248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.183085Z digest=sha256:7cb8b458d747ece4fff1f8e51728fde52cce8b7335ef4c3f664d1078b4a4fab0

Observation 7534699e-44c3-4777-affd-0503001fe5ef · outbound

This paper cites We should introduce a mechanism to reinforce learning during critical times while undermining actions leading to unfavorable outcomes.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation We should introduce a mechanism to reinforce learning during critical times while undermining actions leading to unfavorable outcomes

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:45.259951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.329697Z digest=sha256:79b49538734fdc5c756f2eb4ed18e467da2f3967d148d2fe0c33ed98cf07c96b

Observation 3534af9b-9eb9-495c-aef9-08c1edb9c65d · outbound

This paper cites The policy should focus on maintaining positional advantage to intercept the ball efficiently.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation The policy should focus on maintaining positional advantage to intercept the ball efficiently

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:44.902909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.439086Z digest=sha256:f0ee21fb5d404969988a6df3da5fc94f65aa5bd68b34d3cec7c212bcadc57223

Observation 6cc0f9ff-df86-4f86-b4db-b42a43767c48 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:44.549204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.611895Z digest=sha256:22be3009e0aff5531d1bc8bd42ad9a158e1ff473ca884f6bc84e0b545a658de4

Observation 93396c86-d22d-4fff-aa84-14918688c819 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:44.293375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.728655Z digest=sha256:452392095b81ee36d59cd5987ed5a3bb15b773db3c2bca3bbf7008a6b64d59d2

Observation 586c101f-69b3-4154-81a5-a5d5b6186390 · outbound

This paper cites Implementing a mechanism that prioritizes actions based on proximity to the ball in the x-coordinates can significantly improve action selection during critical moments.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Implementing a mechanism that prioritizes actions based on proximity to the ball in the x-coordinates can significantly improve action selection during critical moments

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:53:43.976039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.854397Z digest=sha256:e1863a2b77c10af95df0f15b79dce5e9017a5b87223c37b3b6540706dcaf7157

Observation d72a2abc-3c3f-4519-8eb0-cd57a5fa90e9 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:43.797107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:41.987874Z digest=sha256:1bb7d41b85bfaa4ade7bbfb6aa9700414fe7840b330bbbf738d39b0512e7f311

Observation 38ef5141-ee5b-463d-9b99-8f1a0e6391c9 · outbound

This paper cites an unresolved cited work.

LLM-Guided Reinforcement Learning: Addressing Training Bottlenecks through Policy Modulation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:53:43.615588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:53:42.129182Z digest=sha256:1e604a7006f33412f2a3a9cedb27022fa15957304630badef70b05f940aa3a97

Pith citing papers

No inbound Pith citation observations are available.