Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:13:32.534554Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2608.03606.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T16:13:32.534554Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
17 of 17 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation afc4603b-355a-40c4-8a58-d983a7be7eac · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents When does return-conditioned supervised learning work for offline reinforcement learning? In Advances in Neural Information Processing Systems (NeurIPS), 2022
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 38f39604-1bfa-4b84-aa82-291ad7815dc3 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Trialbench: Multi-modal ai-ready datasets for clinical trial prediction
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5da2eff5-4b16-4c94-b88a-ef16e1155ff8 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Decision transformer: Reinforcement learning via sequence modeling
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1622e00c-6655-4a91-84e2-7a510dd0df6a · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents RvS : What is essential for offline RL via supervised learning? In International Conference on Learning Representations (ICLR), 2022
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 579859c7-4d9c-472b-9f29-69d511b9bfc7 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents M., and Sun, J
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f59a5af0-eaa2-41da-8b69-569bb770be4a · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Biomni: A general-purpose biomedical ai agent
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ca1d33fd-607d-4ac2-afec-17cb1960ada6 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Offline reinforcement learning as one big sequence modeling problem
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b003f1de-35c9-4997-9668-436c8bb5d323 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents S., Chen, F., Gong, C., Bracken-Clarke, D., Xue, E., Yang, Y., Sun, J., and Lu, Z
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6c72ebba-9545-4f59-9141-343fe579b2ea · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Offline Reinforcement Learning with Implicit Q-Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 944f50eb-2561-49dc-a84b-812052f33c30 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31346935-d93e-4eef-a9ee-0716b4e983db · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c652392-baf9-4b45-99d8-a23efea3cc03 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29627c24-ad87-4a42-975e-dc0a530d7dd2 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7a2ef60-d035-4892-8844-b40f0b341648 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents F., Maximo, M
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4dd17814-fe04-48aa-955f-5411bc480670 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04eae1dd-7bdd-4e88-8e1d-141492700fc3 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents and Kim, Y
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b935df0-b80e-450c-81b0-cf1d5b38c049 · outbound
Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents T., Reed, S., Shahriari, B., Siegel, N., Merel, J., Gulcehre, C., Heess, N., and de Freitas, N
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.