Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:28:49.005957Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2507.10914.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:28:49.005957Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7bbbb9d1-ebfe-40e5-bfc5-0ad85bda4532 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 38e30b3e-872a-4594-a97b-2d85dd73e694 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kakade, and Karan Singh
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c22ab43e-6208-49f9-b1dc-b1931389dcca · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation aff5f864-1eb9-43af-b594-bcd1c0bd90fd · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization On the model-based stochastic value gradient for continuous reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 53d9cd46-22d8-41df-b8f0-1fb39d5ae242 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Infinite-horizon policy-gradient estimation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e4b7239a-3377-4d62-9c1c-a8916a0a3f8e · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization A survey of iterative learning control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b577cd09-eb72-4417-b0a6-512246ff2824 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Difftune: Autotuning through autodifferentiation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ae67067d-d100-466b-a27c-4fcbfdf4cd72 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Differentiable simulation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d8553229-0c08-4253-bc71-88d7b91af176 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5c5787cc-f19a-4698-888a-776b4bb5798f · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Adaptive Regret for Control of Time-Varying Dynamics
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5834f0b8-a3a7-4312-8e75-2200fea83df7 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73c642e0-97fc-41aa-a307-2740153be11f · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Convex Optimization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 77b70f49-23f9-48ea-94cb-9f2a4e0ca9ac · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Control
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cfb56cd-6617-4103-9c81-0fa50984dfa8 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The Nonstochastic Control Problem
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d1cc8004-854f-416c-a8b6-72b72d6cb662 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Ioannou and Jing Sun
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 08d63954-cf62-40d5-a8a4-fb0254a7c730 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Scalable deep reinforcement learning for vision- based robotic manipulation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c583a3a4-abc9-469c-a01e-3e08fe771a03 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kokotovic, and Ioannis Kanel- lakopoulos
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 37184144-6105-4486-9dfe-1e855195dbca · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Harris McClamroch
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0f4c16c9-2acb-4565-8d05-48fd25e588a0 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ea4a7654-2e64-4721-a344-5b33cddb8596 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9d695f52-5c9f-4a72-b53e-6fd09c287e7b · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Universal adaptive control of nonlinear systems
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e5904d1f-e93d-4e61-b8bd-9c8e72e77f23 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2cc030d0-5a9a-464e-b021-2255c0064be2 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Simple random search of static linear policies is competitive for reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2e280188-fd49-4dd3-81d9-0d0ee802a528 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SymForce: Symbolic Com- putation and Code Generation for Robotics
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3eacd447-915c-4609-9a7b-ac126b825d24 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Minimum snap trajectory generation and control for quadrotors
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1dde1e08-914a-4afc-8bb2-9a5ee6373e94 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c6c51d56-7b6d-4a9d-89f4-5b90135c4d05 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Pods: Policy op- timization via differentiable simulation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ee19ca7c-0564-4b71-aa93-20579afcc2a4 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Neural-fly enables rapid learning for agile flight in strong winds
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5acaeffb-f881-4f28-8c20-2c6f66a1c057 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Policy gradient for continuing tasks in dis- counted Markov decision processes
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 58efebf5-1364-46d0-a93a-113be0316e0f · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Wolfgang H ¨onig, Gaurav S
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6caba319-dd72-4866-b2fe-179898315a63 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SPNets: Differentiable Fluid Dynamics for Deep Neural Networks
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 16190867-7e1b-4c16-a4c8-a8e22c11461a · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Proximal Policy Optimization Algorithms
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98fceee1-c4fb-4cf0-add8-3175e8341357 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Parameter-exploring policy gradients
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a577770e-1d7a-4b6c-9662-c6c377d82bb2 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Deterministic policy gradient algorithms
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 84b807a0-1f22-4484-a30b-1502b9061660 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Im- proper Learning for Non-Stochastic Control
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 95620090-b966-4d60-b277-3ef2d0878c0c · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Slotine and W
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 995c03e5-5e44-42f4-bd2f-de62dbb8c435 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8740c015-9516-4fb9-9d42-551529da676b · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 95456459-8247-4628-a495-1f15e4ab6dc6 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Williams
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fadda9a0-c1d0-46fb-986e-77ffa63e60a1 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 698615a6-472b-4027-afc8-8806fafb90ce · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 62a69f1d-b27c-427b-94b6-b8ed1ee28667 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2016b084-451e-4e5b-8ca3-cb7fff685b7b · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 49bd6700-1043-4dd8-a784-46dbf593ee1e · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Note that [vd t ]y = 0 for all desired trajectories
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 19a53eca-05fb-47c8-960b-c5bf7cd92319 · outbound
Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The regularization weights were chosen empirically to be as small as possible while suppressing oscillations
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
No inbound Pith citation observations are available.