Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T12:46:51.707117Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 0 inbound Pith citation observations for arXiv:2606.00680.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T12:46:51.707117Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
11 of 11 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a629de3d-813f-4809-87a8-2afad9e924dc · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Multi-Agent Deep Reinforcement Learning for Liquidation Strategy Analysis
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 708f8489-5155-43e2-8091-ca5256625e5a · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c5fce21-f63a-4560-a4aa-de5af12edb14 · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Soft-Robust Algorithms for Batch Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0cc9998-7849-4936-8731-afdee212879c · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a3d9b67-0648-45cc-abcf-b3944efd32a4 · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief HC” denotes the “HalfCheetah
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d94b858-ee5a-41b2-8a96-06f786e65bac · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief We provide a brief overview of the task below
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dda4bdf0-2a28-435e-b842-2f9e7fc58c6a · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6626df48-ed06-4039-a9e1-05d6a92be0a1 · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Pre-Training for Robots: Offline RL Enables Learning New Tasks from a Handful of Trials
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9168516d-d94a-4c36-b044-4371cee555f1 · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief PhyB achieves superior performance on 8 out of 12 benchmarks and delivers competitive results on the remaining
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38a1fbb3-aa9f-457a-b00e-8315243be804 · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Robust Regularized Policy Iteration under Transition Uncertainty
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34192b9c-ff45-4ec7-abeb-75c0da507b6f · outbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.