Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:47.489929Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2505.19923.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:47.489929Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 893f724a-cc24-4fe6-b9fd-1cd62d354734 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL The mean-wise best results among algorithms are highlighted in bold
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7dfa7fc1-3245-47dd-9d6d-545b04214591 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL RvS: What is Essential for Offline RL via Supervised Learning?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58fe6bdf-e92c-4148-80d2-f214ac4ea121 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf2ca248-7508-4b23-8b16-305983013a7a · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Planning with Diffusion for Flexible Behavior Synthesis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 771cf11e-2c94-4336-a504-57a71276bbda · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Offline Reinforcement Learning with Implicit Q-Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56770b7a-9f6a-4c27-a856-48e770a0cb14 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Reward-Consistent Dynamics Models are Strongly Generalizable for Offline Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e633d48-ad30-4a98-b584-8c13f7e19791 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL S., Ghadirzadeh, A., Chen, X., and Finn, C
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b5041e9-8924-418a-a0bb-aa938abaf12b · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4abe55e6-075f-41d3-b5de-45378bc73e6d · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Q-Ensemble for Offline RL: Don't Scale the Ensemble, Scale the Batch Size
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f965292-2506-4043-9daa-6e2b00cb5c13 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Hyperparameter Selection for Offline Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dbaed76-6543-4524-a810-85d8bc47c469 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a33a587-91c8-495d-a991-4a2389b37c28 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Diffusion Policies as an Expressive Policy Class for Offline Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5795690e-4847-4996-ba18-6a6adf07d008 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d018c99c-319e-49f4-b0e0-a31927ab0fdf · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Policy Expansion for Bridging Offline-to-Online Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cd06e8b-3fc1-415c-8cfe-fe4537e8cdc8 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL ENOTO: Improving Offline-to-Online Reinforcement Learning with Q-Ensembles
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9772347e-cd21-4d34-833d-ee2ba45afc40 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Adaptive Behavior Cloning Regularization for Stable Offline-to-Online Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e191725b-0487-48e2-a856-c8c480519051 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d17ed07-74aa-4a4a-80e6-e2d3388bac09 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Recent works focus on explicit policy constraints for stochastic policies (Wu et al., 2022; Nair et al.,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 68cea97c-3af2-4c50-b1e8-c77a93ac0da6 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 521b3951-9399-427c-96d8-c02efa20145c · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Moreover, the coefficients are updated by maximizing Q-values, lacking the interpretability offered by our method
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cd96905c-98d3-4dec-860b-916782727a32 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41d6a292-3654-4631-86a1-08c3242581f0 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Off-policy deep reinforcement learning without exploration
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 202fe6ac-af03-4756-8236-36e4ef3b38c3 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Extreme Q-Learning: MaxEnt RL without Entropy
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05b7370f-d2b3-493b-ab10-8219106b87b3 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Various CQL variants adjust the constraints or modify the regularizer to avoid excessive pessimism (Lyu et al., 2022; Nakamoto et al., 2024; Mao et al., 2024; Yu et al., 2021)
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e614c14e-8003-4cc3-a08d-0ac9dc9f223a · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Pessimistic Bootstrapping for Uncertainty-Driven Offline Reinforcement Learning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ca01919-e7c7-4a62-9b55-695b0f199b9b · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Efficient Online Reinforcement Learning with Offline Data
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 563b4579-5f45-48bf-9108-6c2930b0c1e2 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Improving TD3-BC: Relaxed Policy Constraint for Offline Learning and Stable Online Fine-Tuning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c8ab753-1a18-465f-a148-48665a553261 · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28a30096-fa6c-4025-92f7-9aca61919fcd · outbound
Learning to Trust Bellman Updates: Selective State-Adaptive Regularization for Offline RL Unresolved cited work
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.