Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:34:23.413516Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2507.22640.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:34:23.413516Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 47d0afb2-e201-4c9c-91b8-555d6cd64d96 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Reinforcement Learning: An Introduction,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 774554ce-b087-4f03-8bc8-5edd85b24f8c · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction From automated to autonomous process operations,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c1f6956-1bd8-4db8-b9ec-41dfddbcc373 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Concrete Problems in AI Safety
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 508c8ec1-2f66-422f-a823-a991a2c02b1b · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Optimal grade transition for polyethylene reactors via NCO tracking,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5c74b45-b2e3-4584-8979-918730fc3169 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Iterative learning control-based batch process control technique for integrated control of end product properties and transient profiles of process variables,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec392a86-50b6-4a89-922f-37f1aa9bb745 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Integrated scheduling and dynamic optimization of grade transitions for a continuous polymerization reactor,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99ea38c8-7d60-44dc-b15a-e78182ec8bf7 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction The general problem of the stability of motion,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37dbd767-c63b-4c32-a7f4-cfbedc629599 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb5cd566-8ea5-4177-b7a8-d22666238a99 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Input Convex Neural Networks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeeaa0b7-fca2-4bff-938c-093e73a7bd10 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Safe Model-based Reinforcement Learning with Stability Guarantees
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc70046b-ce11-4f08-afec-de8e3b61fd2f · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Control Barrier Function Based Quadratic Pro- grams for Safety Critical Systems,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a584b792-8643-44bb-8e73-9c5b521064e9 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Safe and Stable RL (S2RL) Driving Policies Using Control Barrier and Control Lyapunov Functions,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6fbb22bf-b6ad-4ad6-a4a7-bd1ad960f3de · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Constrained Policy Optimization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d6583d2-0b5b-4715-98a1-3576e07c9e16 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Safe Exploration in Continuous Action Spaces
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10db63f3-0676-4291-8c36-01bea85205dc · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Conservative Q-Learning for Offline Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c007703d-cd76-4faa-97ee-60b7af00f21f · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Offline Reinforcement Learning with Implicit Q-Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf93d61e-21ef-4c72-b842-e954f85b855f · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction MOPO: Model-based Offline Policy Optimization
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ccfd0f9-726d-4af3-afb6-4699c5e56739 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Actor–Critic Physics-Informed Neural Lyapunov Con- trol,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 96778281-0ef6-4a6c-8dc0-c64a2fd2973b · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Distributional Reinforcement Learning with Quantile Regression
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 370834ac-9400-41f1-a108-4603f8744f26 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction EKG-AC: A New Paradigm for Process Indus- trial Optimization Based on Offline Reinforcement Learning With Expert Knowledge Guidance,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3826ad44-7d6d-4395-9796-201a044176c5 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Optimal Control Via Neural Networks: A Convex Approach
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feab1efd-3c40-4dfb-9e6a-13241dd895b7 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Differentiable Convex Optimization Layers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb8b85ae-3c1d-4330-9859-32c9be6729d1 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction OptNet: Differentiable Optimization as a Layer in Neural Networks,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d8fd4b8-b759-4fe9-89ab-822f849d849d · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Polymer grade transition control using advanced real-time optimization software,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f69c0586-8e19-4c5e-bd61-6aebe21db289 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Polymer grade transition control via reinforcement learning trained with a physically consistent memory sequence-to-sequence digital twin,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0ba9b1af-a386-4ca0-b2af-30e290e13a68 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction A benchmark environment motivated by industrial control problems,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e3f99c7-1864-46b0-97c3-b30cd9b1c067 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef465ad6-a07f-4c36-8315-f2958d7f4a43 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction PC-Gym: Benchmark Environments For Process Control Problems
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb487b2f-da8a-49e3-86a9-7867a95ee095 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction End-to-End Safe Reinforcement Learning through Barrier Functions for Safety-Critical Continuous Control Tasks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 203fd197-31c7-47a0-99e1-e90cc81eff2c · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Offline reinforcement learning methods for real-world problems,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dae50e37-7c3d-4137-9e10-976a79f04421 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe6784ac-7e2f-4dcc-b52e-8fa68cd782bf · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction A survey on offline reinforcement learning: Taxonomy, review, and open problems,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dce9edd8-9547-4c5a-8e76-9faad7a9eafd · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Stabilizing off-policy q-learning via bootstrapping error reduction,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf7f5e39-0013-46cb-9550-a3bcbe6650ab · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Deep Reinforcement Learning with Double Q-learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f22b21aa-dcd3-4daa-a3f5-ab7a82bb7621 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Human-level control through deep reinforcement learning,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d55ea92-d8cd-437c-b857-b4e5247f3f56 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Behavior Regularized Offline Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 176e6c8e-b993-44a6-a594-cc0c91212bb0 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09431ac7-47f5-49c0-a3bc-9ec13e86c089 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Comparative Study of Machine Learning and System Identification for Process Systems Engineering Dynamics,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc6fd994-6286-4aed-8703-3ba5690351f0 · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction Polymerization reactor control using autoregressive-plus Volterra- based MPC,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8ff2815d-9103-4710-be2f-c89055b741cc · outbound
Safe Deployment of Offline Reinforcement Learning via Input Convex Action Correction OptNet: Differentiable Optimization as a Layer in Neural Networks
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.