Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:12:07.125494Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2606.31320.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:12:07.125494Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5c4f1ad0-9cb6-446b-bef9-689b86ce0c60 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Control Barrier Functions: Theory and Applications
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7f051ca2-e846-4547-8969-46da3a60e246 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition doi: 10.1145/3744351
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 38be1716-538c-4c67-aee2-ab0b15a0d9bc · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Lee, Matthew Tan, Yuke Zhu, and Jeannette Bohg
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ceb67701-16a4-4252-8dda-dade35141081 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Learning to Walk in the Real World with Minimal Human Effort
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8615bfd7-5425-4501-bc16-d0f9940f54e9 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Annual Review of Control, Robotics, and Autonomous Systems7(2024).https://doi.org/10.1146/ANNUREV-CONTROL-071723-102940
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 18af1742-2420-49d3-b5e6-59d05959123d · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Ibarz, J
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cb52c102-f21b-463e-8509-e94104ead8ff · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition cc/paper/2021/hash/85ea6fd7a2ca3960d0cf5201933ac998-Abstract.html
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 210dbb4f-da32-40ff-ba46-b8ffdd8c9943 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Provably Safe Reinforcement Learning: Conceptual Analysis, Survey, and Benchmarking
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 20ef1545-8009-4f72-abe6-15934da4e59f · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Safe Reinforcement Learning using Action Projection: Safeguard the Policy or the Environment?
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ee1fb851-bcdc-4668-9f5c-065fe208ec7b · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition arXiv preprint arXiv:2509.21014 (2025)
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation da4d45d8-3c5f-4df3-8dd0-cdea7da1886a · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition In: NASA Formal Methods Symposium
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 637a075c-bb27-4f0c-89b6-dc5bcad91686 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Benchmarking Batch Deep Reinforcement Learning Algorithms
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 22ee596a-adbc-427e-9425-0243d819bfc2 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Proximal Policy Optimization Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5a60eb90-bd82-4a9d-ad67-aaf7218c74dd · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d106a98a-4357-401f-8c5a-21a19dc16afe · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Reinforcement Learning with Adaptive Regularization for Safe Control of Critical Systems
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 66b11d40-83c4-4d4a-8b22-81ff3d9dd91e · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Linear model predictive safety certification for learning-based control
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 582fa6b4-6588-4cf6-910e-19f539cd82cf · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition ISBN 978-1-4503-1996-6
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c3aa37ef-0bb5-41e8-8925-9293bbb4cc38 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Linhai Xie, Sen Wang, Stefano Rosa, Andrew Markham, and Niki Trigoni
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 498762bc-ae7e-41fc-b637-80a375a18c4c · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Stable and Safe Reinforcement Learning via a Barrier-Lyapunov Actor-Critic Approach
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c28926c-b38e-48d9-8234-37adc57e3e17 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition That is, there exist b(s)∈Randg(s)∈Rm such that ˆ∆(s,a) =b(s) +g(s) ⊤a+εlin(s,a),|εlin(s,a)|≤ϵlin(s)(22) for all actionsaon this segment
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cae868ec-ab63-4c34-b105-f53be0792bca · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d73eea61-3bfc-4ced-ad97-6aa64ca6c260 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b6865a8c-3410-4022-896d-835c977c9034 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition qCMdcjpSuxuYQzuZyBqqjO9S8DY=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e45dc118-bb26-487b-b5ea-27586c9bb944 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 56d335bd-1817-4bf2-a3f2-c1197373fce8 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition The system model can be found at (Tian et al., 2024)
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ae0de3cb-6074-4832-b6ea-d2fe5963d2c6 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition In our case study, we set the initial position of the quadrotor ass0xyz ={1.5,1.5,1.5}and the target position of the quadrotor asˆsxyz ={2.5,2.5,2.5}
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b7ae71bf-47ee-4e78-9539-9526fc9cf917 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition We observe that training of theSimplexis graduallydivergingwitha largecriticloss, asshown inFig.11
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2838b799-7e82-4d6b-97e1-696d672721dc · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2e692e5f-97aa-449a-ba4a-11c885d7b2b0 · outbound
Safe Online Learning via Smooth Safety-Structured Policy Composition For simple tasks, such as cartpole and glucose, the agent could learn using the data generated by the safe policy
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.