Pith. sign in

Paper Citation Record · LEDGER

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization

As of 22 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2507.10914.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.10914 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:28:49.005957Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy41
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7bbbb9d1-ebfe-40e5-bfc5-0ad85bda4532 · outbound

This paper cites Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Optimal Algorithms for Online Convex Optimization with Multi- Point Bandit Feedback

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.939978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:44.566036Z digest=sha256:5ed899605674c2ce5aff39444940b07dff937228fa617c5031da2af24f2dbc83

Observation 38e30b3e-872a-4594-a97b-2d85dd73e694 · outbound

This paper cites Kakade, and Karan Singh.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kakade, and Karan Singh

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.783691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:44.642567Z digest=sha256:868a38afadbb07b74b8c7aa7f59785a94c13ecc39d66561ce1793644a5ed78b8

Observation c22ab43e-6208-49f9-b1dc-b1931389dcca · outbound

This paper cites Zico Kolter.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.657831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:44.751125Z digest=sha256:f9035f7ff19004f9052ca29fa2017368e6593350b70fb1feddcb8dfa221b579d

Observation aff5f864-1eb9-43af-b594-bcd1c0bd90fd · outbound

This paper cites On the model-based stochastic value gradient for continuous reinforcement learning.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization On the model-based stochastic value gradient for continuous reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.545675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:44.810525Z digest=sha256:3582de627733c995a84a4aab5df146181483c9680293a8ad8398e861cf0f0c9f

Observation 53d9cd46-22d8-41df-b8f0-1fb39d5ae242 · outbound

This paper cites Infinite-horizon policy-gradient estimation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Infinite-horizon policy-gradient estimation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.444040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.028275Z digest=sha256:306e555a57b5beac73bd0b74fafb8661017ef5ae4e2a7083a60d454a4662175b

Observation e4b7239a-3377-4d62-9c1c-a8916a0a3f8e · outbound

This paper cites A survey of iterative learning control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization A survey of iterative learning control

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.300171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.113141Z digest=sha256:d023659aa51a10935320017bc3547a5c832f88028f3ad10583ee70b2c52516ba

Observation b577cd09-eb72-4417-b0a6-512246ff2824 · outbound

This paper cites Difftune: Autotuning through autodifferentiation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Difftune: Autotuning through autodifferentiation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.184190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.204092Z digest=sha256:d45c313e494386a56945086408746e23b14b486e053478e3310c95bef38b5819

Observation ae67067d-d100-466b-a27c-4fcbfdf4cd72 · outbound

This paper cites Differentiable simulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Differentiable simulation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:55.049415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.322459Z digest=sha256:53bf0f736962b1b4f7078ede11c0da3c9dbdde62d2fef238160407ecc613fd6c

Observation d8553229-0c08-4253-bc71-88d7b91af176 · outbound

This paper cites Zico Kolter.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zico Kolter

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.944338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.418546Z digest=sha256:da4949eb5e51e711aa8e503b33ed03a12fbd200d4f3f0a83581228d55a3e4ccf

Observation 5c5787cc-f19a-4698-888a-776b4bb5798f · outbound

This paper cites Adaptive Regret for Control of Time-Varying Dynamics.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Adaptive Regret for Control of Time-Varying Dynamics

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.791007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.461848Z digest=sha256:648dc8bce50950d4580aae360672ee87eb9b6d016baed2f47f8c15b0ad8b38b5

Observation 5834f0b8-a3a7-4312-8e75-2200fea83df7 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:45.574042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:45.574042Z digest=sha256:802035c0ee4bc1060feb6969514c89136f87ef628c5737d07b2d19068d295d74

Observation 73c642e0-97fc-41aa-a307-2740153be11f · outbound

This paper cites Introduction to Online Convex Optimization.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Convex Optimization

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.677312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.652824Z digest=sha256:a9fbe72527ed9313576967cb8e7997057adbda3f4258102c06978146662ed1cf

Observation 77b70f49-23f9-48ea-94cb-9f2a4e0ca9ac · outbound

This paper cites Introduction to Online Control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Introduction to Online Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:45.738890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:45.738890Z digest=sha256:0e1c328ae8cc81d25d16fed933eac0f8b6c7582e131d88addfd04196493a08ac

Observation 5cfb56cd-6617-4103-9c81-0fa50984dfa8 · outbound

This paper cites The Nonstochastic Control Problem.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The Nonstochastic Control Problem

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.567972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.804696Z digest=sha256:3e98349251615af60f6d392d5bfb63ec31e82c9473893b8b062970b8f6692c12

Observation d1cc8004-854f-416c-a8b6-72b72d6cb662 · outbound

This paper cites Ioannou and Jing Sun.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Ioannou and Jing Sun

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.451430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:45.934723Z digest=sha256:92309aed94e96a01473166da9c1bb2b325e3e68c1b580d7b73806caa734d87fd

Observation 08d63954-cf62-40d5-a8a4-fb0254a7c730 · outbound

This paper cites Scalable deep reinforcement learning for vision- based robotic manipulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Scalable deep reinforcement learning for vision- based robotic manipulation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.314728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.031777Z digest=sha256:6b65b1dc831b11970cf895198ec47565092c38b36ed54c4fe59e9f47a467a249

Observation c583a3a4-abc9-469c-a01e-3e08fe771a03 · outbound

This paper cites Kokotovic, and Ioannis Kanel- lakopoulos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Kokotovic, and Ioannis Kanel- lakopoulos

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.190188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.081046Z digest=sha256:c17c692dbafafc7a67eae43868370e5605aba0f9d103069d1b2b8189af0a8be5

Observation 37184144-6105-4486-9dfe-1e855195dbca · outbound

This paper cites Harris McClamroch.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Harris McClamroch

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:54.077934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.205280Z digest=sha256:964bee5fe615af9ba4d38e965e1ac67af722216ee2fc759a11f1b6a940894770

Observation 0f4c16c9-2acb-4565-8d05-48fd25e588a0 · outbound

This paper cites Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Emile Anand, Yingying Li, Yisong Yue, and Adam Wierman

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.953338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.388185Z digest=sha256:2a5c50ed7e5eec2310b83a3f6934a86bba03e757976ae43cff1f6e335970278d

Observation ea4a7654-2e64-4721-a344-5b33cddb8596 · outbound

This paper cites Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Fengze Xie, Emile Anand, Soon-Jo Chung, Yisong Yue, and Adam Wierman

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.832829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.489774Z digest=sha256:d1f3e253fa2fbc50187b2616fa0acb9778894b9f19f5c3200e1184f419df61f5

Observation 9d695f52-5c9f-4a72-b53e-6fd09c287e7b · outbound

This paper cites Universal adaptive control of nonlinear systems.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Universal adaptive control of nonlinear systems

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.702159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.584372Z digest=sha256:ba624359cb78792cc58027353a954695e48fdf2ab61ba26cc2989118d0faff42

Observation e5904d1f-e93d-4e61-b8bd-9c8e72e77f23 · outbound

This paper cites Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Jedidiah Alindogan, Matthew Anderson, and Soon-Jo Chung

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.596700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.693680Z digest=sha256:6b65cf3359a3a4b84782647af597829588347a270aa2acb966f8da616e7b51ec

Observation 2cc030d0-5a9a-464e-b021-2255c0064be2 · outbound

This paper cites Simple random search of static linear policies is competitive for reinforcement learning.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Simple random search of static linear policies is competitive for reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.433582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.829555Z digest=sha256:004a2e011821b7950067f38e5287dd5ed71338bbe46b66f701fdad695d039874

Observation 2e280188-fd49-4dd3-81d9-0d0ee802a528 · outbound

This paper cites SymForce: Symbolic Com- putation and Code Generation for Robotics.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SymForce: Symbolic Com- putation and Code Generation for Robotics

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.289204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:46.960052Z digest=sha256:b47c69e21be238f12e51395d2c9713e89e33fa7304a0ec7e61523aa3515547a3

Observation 3eacd447-915c-4609-9a7b-ac126b825d24 · outbound

This paper cites Minimum snap trajectory generation and control for quadrotors.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Minimum snap trajectory generation and control for quadrotors

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.108160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.053906Z digest=sha256:cd29cc8b12a3081de5886b1f36eed0e0023486442d0e8106fde6ee1dc2c3f7ac

Observation 1dde1e08-914a-4afc-8bb2-9a5ee6373e94 · outbound

This paper cites Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Nonlinear and adap- tive intelligent control techniques for quadrotor UA V–a survey

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:53.010365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.170933Z digest=sha256:4497c36e4e65d716da8627f4b30677bb0a9efdbb3d69259e59cfbb728edc90c7

Observation c6c51d56-7b6d-4a9d-89f4-5b90135c4d05 · outbound

This paper cites Pods: Policy op- timization via differentiable simulation.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Pods: Policy op- timization via differentiable simulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.833150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.277587Z digest=sha256:4aa6516c87da9d784dcafb7222c6f2c737233f03766452158b60493531c41459

Observation ee19ca7c-0564-4b71-aa93-20579afcc2a4 · outbound

This paper cites Neural-fly enables rapid learning for agile flight in strong winds.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Neural-fly enables rapid learning for agile flight in strong winds

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.646712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.374923Z digest=sha256:aac62f660afd41c01b697b523f5a8e2423b58fa2737009d42c7885b08564df98

Observation 5acaeffb-f881-4f28-8c20-2c6f66a1c057 · outbound

This paper cites Policy gradient for continuing tasks in dis- counted Markov decision processes.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Policy gradient for continuing tasks in dis- counted Markov decision processes

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.452974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.469258Z digest=sha256:58e4e4a6c9abee11130eaf645495597402ea523b48d797889938356a7f1054a8

Observation 58efebf5-1364-46d0-a93a-113be0316e0f · outbound

This paper cites Preiss, Wolfgang H ¨onig, Gaurav S.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Preiss, Wolfgang H ¨onig, Gaurav S

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:52.270065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.569127Z digest=sha256:ceb24fca0f5d90198ed9fbba613f240988dbebefa78c5f76543b5807ae613c33

Observation 6caba319-dd72-4866-b2fe-179898315a63 · outbound

This paper cites SPNets: Differentiable Fluid Dynamics for Deep Neural Networks.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization SPNets: Differentiable Fluid Dynamics for Deep Neural Networks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.939053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.694501Z digest=sha256:66d582d71a135aeb2093ce0d3476d6c6e21e0639ea0298d337de62ad6bb5793e

Observation 16190867-7e1b-4c16-a4c8-a8e22c11461a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:28:47.770195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:28:47.770195Z digest=sha256:ab23d5051a2d2f7e039d61efedaa3df6409da043bf7e4e2c0e6eca1754ee5b5c

Observation 98fceee1-c4fb-4cf0-add8-3175e8341357 · outbound

This paper cites Parameter-exploring policy gradients.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Parameter-exploring policy gradients

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.713345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:47.896497Z digest=sha256:bae5159fd9c28765590ccc908a0f0fa466d753c342a6c33cc445d5ff55bc9aa9

Observation a577770e-1d7a-4b6c-9662-c6c377d82bb2 · outbound

This paper cites Deterministic policy gradient algorithms.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Deterministic policy gradient algorithms

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.535018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.002435Z digest=sha256:d96a1d3d92b4e46f8a585250cb6a5ab97d12473a3a0e8f688fa3a8330cd8092c

Observation 84b807a0-1f22-4484-a30b-1502b9061660 · outbound

This paper cites Im- proper Learning for Non-Stochastic Control.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Im- proper Learning for Non-Stochastic Control

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.399197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.086917Z digest=sha256:8fc58e784e079fad0a318a169f0e55c8b577ee7d40158755b8703cddc3c5505e

Observation 95620090-b966-4d60-b277-3ef2d0878c0c · outbound

This paper cites Slotine and W.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Slotine and W

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.256899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.154256Z digest=sha256:d7b155bd6281698e76ad98e595a23853dc3ca726b7f0852a69ce7551c2e6c8db

Observation 995c03e5-5e44-42f4-bd2f-de62dbb8c435 · outbound

This paper cites Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Do differentiable simulators give better policy gradients? In International Conference on Ma- chine Learning (ICML) , 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:51.073789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.280899Z digest=sha256:c308e40f5a4d5132a60f355955c4d7f39c35e05598064359bf13ccd9bb524cca

Observation 8740c015-9516-4fb9-9d42-551529da676b · outbound

This paper cites an unresolved cited work.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-06T17:28:50.824812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.364930Z digest=sha256:86d3925b59f2761308a6298152d3e9a95e9fbdecf31ceda54352ded14db4e367

Observation 95456459-8247-4628-a495-1f15e4ab6dc6 · outbound

This paper cites Williams.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Williams

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.611937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.485867Z digest=sha256:bbe6822b05fa76fdfd3c61d39d8499095a36b1eb740ca0643bce4cce5025c52a

Observation fadda9a0-c1d0-46fb-986e-77ffa63e60a1 · outbound

This paper cites JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization JAX- FEM: A differentiable GPU-accelerated 3D finite ele- ment solver for automatic inverse design and mechanistic data science

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.443986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.576728Z digest=sha256:1af85cd09e500f03787fbc72754a2ee23f02566040bb6f2ccacc6cfca8355a67

Observation 698615a6-472b-4027-afc8-8806fafb90ce · outbound

This paper cites Zavlanos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:50.162035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.653726Z digest=sha256:24f5c2b0ae6cc74fab06098d3609efa3d2384044bdcdcbcf763b9371b025521e

Observation 62a69f1d-b27c-427b-94b6-b8ed1ee28667 · outbound

This paper cites Zavlanos.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Zavlanos

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.969562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.767239Z digest=sha256:b12c91cf642507867f938c6b9d21f180019fc9cf3da9e59bad0e2856876fcc67

Observation 2016b084-451e-4e5b-8ca3-cb7fff685b7b · outbound

This paper cites The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The system state comprises the car’s position p ∈ R2, body-frame velocity v ∈ R2, heading angle r ∈ so(2), angular velocity ω ∈ so(2), and steering angle ψ ∈ R

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.690094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.862950Z digest=sha256:5487148fd02509754d4c2d44893f2e317365ba48c2e3fd02186c2ca312abe14d

Observation 49bd6700-1043-4dd8-a784-46dbf593ee1e · outbound

This paper cites Note that [vd t ]y = 0 for all desired trajectories.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization Note that [vd t ]y = 0 for all desired trajectories

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.531718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:48.927624Z digest=sha256:db2b2ee2a33d1c8af7822555a3a735669dc2f627c1d9879f2a9bad25c040384b

Observation 19a53eca-05fb-47c8-960b-c5bf7cd92319 · outbound

This paper cites The regularization weights were chosen empirically to be as small as possible while suppressing oscillations.

Fast Non-Episodic Adaptive Tuning of Robot Controllers with Online Policy Optimization The regularization weights were chosen empirically to be as small as possible while suppressing oscillations

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:28:49.257124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T17:28:49.005957Z digest=sha256:e9da8b8e2dcd78980bdccf6cbd0c53f9f2402bc471eda8fe1e169147b5c9fc4f

Pith citing papers

No inbound Pith citation observations are available.