Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T17:52:45.955765Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2608.03502.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T17:52:45.955765Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation eca521a5-94e8-42c0-bdea-477d7231d8d0 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8c52a268-6028-4ca4-829d-535ce9648ece · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Human -level control through deep reinforcement learning,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a99d210d-1c5b-4a59-9226-4389fc8a1654 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Proximal Policy Optimization Algorithms
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b9051e0-2a97-45c1-b01d-072f5380c34e · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Soft Actor-Critic: Off-Policy Maximum Entropy Deep RL,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d843c560-c625-4da2-bc14-921e0372db2b · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Reinforcement Learning Unplugged: Benchmarks for Offline RL,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dd13bf29-602d-49ef-9f3d-a4e725215c03 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9babcc4e-2d69-4e2c-808b-8871c7ecbfe7 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Reflexion: Language Agents with Verbal Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f5321e7-3c2b-4115-a645-a968551162ec · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Toolformer: Language Models Can Teach Themselves to Use Tools
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59738885-9cdd-4ab7-bd73-aff76588b9b8 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cba9dc8f-745c-4e70-8b2a-0763c6f61683 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks APCodec: A Neural Audio Codec with Parallel Amplitude and Phase Spectrum Encoding and Decoding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63ef0a9b-0eb2-4062-acbb-75db8b113e97 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks ReST meets ReAct: Self-Improvement for Multi-Step Reasoning LLM Agent
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5bf6165-612b-4c71-b364-73539460400a · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Language-Conditioned Reinforcement Learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a961a01-c5c1-4357-8ec0-85ef8cbcc08f · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks PreRoutGNN for Timing Prediction with Order Preserving Partition: Global Circuit Pre-training, Local Delay Learning and Attentional Cell Modeling
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fdd51f49-66ea-4dd4-944f-60bc90a8f274 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks LLM-Guided Symbolic Planning,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5aa58d80-a10b-4a1a-9b01-a8b2cf53862d · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Improving Reproducibility in Reinforcement Learning Research,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f327194e-3b15-4708-98e3-97f18730a82f · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Representation Learning: A Review and New Perspectives,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0ba9a4bf-c66f-44a4-8a4c-e10d85df9e8d · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Courville, I
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a3b46684-97ba-482e-8070-bcf62c816258 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks A Geometric Perspective on Optimal Representations for Reinforcement Learning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 85594ee2-b6db-4aa6-9fc7-c82b88b4819a · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Deep Reinforcement Learning that Matters,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb977f77-e46a-4a77-96c1-f02aaaaf45c1 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7af8bce-84f6-4ffd-9414-2780caaabab5 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fecb19fb-45d1-49b8-805d-fced83feb892 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 58dd0910-0449-4b33-9bfa-989fa0fc5661 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4fa53196-163c-4484-b91c-9a8bdab40386 · outbound
Hybrid LLM-Augmented Reinforcement Learning Agents for Complex Sequential Decision Tasks Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.