Pith. sign in

Paper Citation Record · LEDGER

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents

As of 7 August 2026, this Paper Citation Record lists 100 of 197 outbound references and 2 inbound Pith citation observations for arXiv:2507.13491.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13491 v1

Coverage vector

measured 100 of 197 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:28:55.851695Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T02:14:58.076642Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T01:17:30.887576Z

Reference resolution

100 of 197 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved96
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e7c15d3a-7095-4920-9d5d-3d6c1593fed2 · outbound

This paper cites Preference-Based Policy Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Preference-Based Policy Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.313392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.313392Z digest=sha256:1ddf5487ac003838933f74570c088c6d3ddfdcefd7a8b7d1091a1dbe1d45db2b

Observation 736f3e60-5c53-4c76-959e-8599b463cd58 · outbound

This paper cites Ames, Samuel Coogan, Magnus Egerstedt, Gennaro Notomista, Koushil Sreenath, and Paulo Tabuada.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Ames, Samuel Coogan, Magnus Egerstedt, Gennaro Notomista, Koushil Sreenath, and Paulo Tabuada

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.348210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.348210Z digest=sha256:71c228ae16a69550fe5e0679096f2cff43c3346e252a0ab3517e6c682b640ca8

Observation f9df93c7-572c-43c2-a873-0318a593964b · outbound

This paper cites Concrete Problems in AI Safety.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Concrete Problems in AI Safety

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.452672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.452672Z digest=sha256:62404ba2d5b54217d023bfd1135e5c396009dd4db7c8058b89215c2351204ff8

Observation efb99a98-d436-42d6-9235-86b8f095e65f · outbound

This paper cites Zico Kolter.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Zico Kolter

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.519797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.519797Z digest=sha256:49217fe4a519019359f05688987c13f06e3744f8d33f087f0b7e5ce58ad2272c

Observation 279dd837-77d0-419d-9e8a-9a0a0b40d59f · outbound

This paper cites A Painless Deterministic Policy Gradient Method for Learning-based MPC.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents A Painless Deterministic Policy Gradient Method for Learning-based MPC

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.620576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.620576Z digest=sha256:9f70b456545145c1cf1a11cd753962e922a0d98a3688c84663de8a06b4a16f3e

Observation 74523618-bc9c-4258-8f83-4476fcbdb7fc · outbound

This paper cites Andrychowicz et al.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Andrychowicz et al

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.695472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.695472Z digest=sha256:5ffc7c7af9257f6272a99e538091e443a2d49c91bd1af811c71a39a83b695c7d

Observation 7df2d093-7db2-4c00-aef9-79947ba475b5 · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.789971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.789971Z digest=sha256:1c7756161e662199b9cdf42c1e7e0b7be22153ae13aeabf4995ff239ffde5359

Observation c6c96e72-488e-406e-b02b-cbd403035d0b · outbound

This paper cites MPC-based reinforcement learning for economic problems with application to battery storage.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents MPC-based reinforcement learning for economic problems with application to battery storage

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.844883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.844883Z digest=sha256:c6817ca26796276764baef4d42f6900db3fc57e0bb79352800eb8ea8c389f2fd

Observation 0639dd66-ad08-4362-9d8e-7181577964a9 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:44.984676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:44.984676Z digest=sha256:d5dbd7eec2f04b3b9100597826d1a03da44339bf98c0b75c391ff952efde9d7f

Observation ee64041e-14be-4679-915b-e94e3e37f4e0 · outbound

This paper cites Local-Global Learning of Interpretable Control Policies: The Interface between MPC and Reinforcement Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Local-Global Learning of Interpretable Control Policies: The Interface between MPC and Reinforcement Learning

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T16:29:09.504750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:28:45.059769Z digest=sha256:a5608a38a072ff69d33ba1054dd67334475d06db55a372d2e37395a00a15ee09

Observation e10bdf5b-55f9-4429-8028-5a0167c0c045 · outbound

This paper cites Gradient-Based Framework for Bilevel Optimization of Black-Box Functions: Synergizing Model-Free Reinforcement Learning and Implicit Function Differentiation.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Gradient-Based Framework for Bilevel Optimization of Black-Box Functions: Synergizing Model-Free Reinforcement Learning and Implicit Function Differentiation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.136798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.136798Z digest=sha256:9a7bbf0697cbdbce70edb1a05bdf3a618127d72deb8f8989f75b0e1c37d96861

Observation 7e45a85b-b174-41cd-b8ba-18859ed100df · outbound

This paper cites Bellemare, Will Dabney, and R´ emi Munos.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bellemare, Will Dabney, and R´ emi Munos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.265231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.265231Z digest=sha256:52f062e1de10fb5f6f5aed03ba5f3a35a83e0dd48ec1686899d8c91296d643c6

Observation d3a9b809-f4c7-4026-8e52-51fe473b0f40 · outbound

This paper cites Bellman and R.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bellman and R

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.332034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.332034Z digest=sha256:2b7c82b80c1284a5076d9a3d58e60107d041268d524b6e772a1c190236d3b6e8

Observation d64caec5-61eb-42b2-a2ca-86a1f1031cb1 · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.439282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.439282Z digest=sha256:d27fb1412e1dcf30314f7ebf62c5a8101cbf0b07444de8cd80fccc749bb60793

Observation 7eec9bcf-9dad-4f14-be9f-c1a51c041c06 · outbound

This paper cites Bellman and Stuart E.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bellman and Stuart E

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.540272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.540272Z digest=sha256:18a66fe9f7f20a5b2b533b3f88bcec34762d777f0876ee41689e97e29ab0ad44

Observation a94bedd2-011e-4ad1-a77d-0967f5ae8ea3 · outbound

This paper cites Schoellig.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Schoellig

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.695771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.695771Z digest=sha256:4d478822f8bede237b96cabdc1ab4819a924cb9d9a628f9be7039249916d28ae

Observation 70d7ae2f-36db-4924-acdd-859280358ef8 · outbound

This paper cites Bertsekas and J.N.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bertsekas and J.N

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.768307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.768307Z digest=sha256:c6fca212d8aa2c41f2ebed99e44e981ea2f8670d84953114e90e3caa93e7656a

Observation e89928d0-be98-41fd-bd19-ef0392e64aaa · outbound

This paper cites Global Optimality Guarantees For Policy Gradient Methods.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Global Optimality Guarantees For Policy Gradient Methods

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:45.850097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:45.850097Z digest=sha256:b7a2095811600789afd4eebdf55ce4463be4fa099dbdce6a205f1e80de18a57b

Observation edd22dfd-71d3-4edb-b8c1-83b28c67ec73 · outbound

This paper cites Differentiable optimization-based control policy with convergence analysis, 2025.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Differentiable optimization-based control policy with convergence analysis, 2025

Reference 19

Resolution
verified exact
raw_fallback, observed 2026-08-06T16:29:09.213959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:28:45.951592Z digest=sha256:b37898c54701790a03f0004245d8769a4d7c31b3e0bbe3247dfdf93c7098ca1a

Observation 44f8d265-bd48-4174-907b-d3713d88d0c8 · outbound

This paper cites A survey on high- dimensional gaussian process modeling with application to Bayesian optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents A survey on high- dimensional gaussian process modeling with application to Bayesian optimization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.050703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.050703Z digest=sha256:82b460f347e18f9a996d1e236d0cc56ae65d5d818712c7d09fcbcce4ba1b1f4f

Observation 102b9f4c-3129-4893-97b9-b00f3e38031e · outbound

This paper cites Blondel and John N.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Blondel and John N

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.156986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.156986Z digest=sha256:2eea32267735d84d5508877c05eaa628e25225ee9dd6a73091984812dc733539

Observation c330312f-fc5c-4212-8054-8f184cb8be85 · outbound

This paper cites Time-varying gaussian process bandit optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Time-varying gaussian process bandit optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.328192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.328192Z digest=sha256:f5c3637ad73b1b2603710c94f8c804c7ffb142de9bff456084746a4ddbf87433

Observation 7916b5ca-9698-422b-881a-7bb00e47feb2 · outbound

This paper cites Optimization of the model predictive control meta-parameters through reinforcement learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Optimization of the model predictive control meta-parameters through reinforcement learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.479989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.479989Z digest=sha256:5c001e126748aee79762e9109694922ce566394e59378b9dc81e74798f8f5dd0

Observation ec0938e2-b854-4aef-bd12-b00dc26d9cd8 · outbound

This paper cites Safe Learning in Robotics: From Learning-Based Control to Safe Reinforcement Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Safe Learning in Robotics: From Learning-Based Control to Safe Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.686707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.686707Z digest=sha256:ad16ff9c1707bdb33118a0290c28355925266c38a7948a1b365a7c0b5ffa4507

Observation 699c5d58-a22e-4094-9152-2cf237b4d36c · outbound

This paper cites On controller tuning with time-varying bayesian optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents On controller tuning with time-varying bayesian optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.823255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.823255Z digest=sha256:651937b703a3b9520b03e600548f516ea4ad3416b888acd3fd1c210f19404b3e

Observation 5fac73f7-f300-4eb5-9c54-2955187f2821 · outbound

This paper cites Reinforcement Learning of the Prediction Horizon in Model Predictive Control.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Reinforcement Learning of the Prediction Horizon in Model Predictive Control

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:46.982388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:46.982388Z digest=sha256:a0576bb56c784fc5d4a518dce4eaefbf35184aa35d6be7be44e688e800eb8442

Observation 887bc6a0-2a96-40e9-96a6-4d7120bcee16 · outbound

This paper cites MPC-based Reinforcement Learning for a Simplified Freight Mission of Autonomous Surface Vehicles.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents MPC-based Reinforcement Learning for a Simplified Freight Mission of Autonomous Surface Vehicles

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:29:08.828872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:28:47.092643Z digest=sha256:3dfca0097ce480a70f7bd09ab2e1ab9e30091922724dc952f7f25a086ca36930

Observation a305b5f0-1d9e-4f0c-b911-0ed010ca3d25 · outbound

This paper cites Chan, Georgios Makrygiorgos, and Ali Mesbah.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Chan, Georgios Makrygiorgos, and Ali Mesbah

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:47.285611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:47.285611Z digest=sha256:ae13a981722b9a75ca7225c1316ac8ec2a1e0ff11b1bc9b9a0864c18549aea69

Observation 9d9d0d33-fdc7-4195-9439-1c98130e8f3b · outbound

This paper cites Chan, Joel A.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Chan, Joel A

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:47.464764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:47.464764Z digest=sha256:74e7df3df39d1b2e008193a933d2d32ac87a15ab018af2a905578fb91d7021ca

Observation d1d7ee7f-01f0-4376-89fc-33052a11e5c3 · outbound

This paper cites Chan, Joel A.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Chan, Joel A

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:47.634740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:47.634740Z digest=sha256:f2c7e3baf1ac8895700e19d90fc7eab78cfac030da9f6d5e3e855a5f34e54caa

Observation d457332e-e506-43d3-acf3-fa455d5f8e69 · outbound

This paper cites Goal- conditioned reinforcement learning with imagined subgoals.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Goal- conditioned reinforcement learning with imagined subgoals

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:47.786612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:47.786612Z digest=sha256:4088f7f24644b6f14c99ec9ebb4c37b1337b40ff7c3059d8808630b5e3349e82

Observation 341b75b4-268e-4423-ad7b-b587ad2b3266 · outbound

This paper cites Gnu-RL: A precocial reinforcement learning solution for building hvac control using a differentiable MPC policy.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Gnu-RL: A precocial reinforcement learning solution for building hvac control using a differentiable MPC policy

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:47.902314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:47.902314Z digest=sha256:1b87ca5ac4fa93021cb631c385edecfb2ec3a410bea868f9a8c6682c6f0ac9a3

Observation 15a8427b-cd1e-4084-a8ce-f952fa813103 · outbound

This paper cites Intrinsically motivated reinforcement learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Intrinsically motivated reinforcement learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.026251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.026251Z digest=sha256:be2b8af9a2d6d7e974247020c826c240740fddd2e4f34f73d10042b86af83079

Observation bd2f87d3-5984-4d34-8c62-8fc395dabd9a · outbound

This paper cites Run- indexed time-varying Bayesian optimization with positional encoding for auto-tuning of controllers: Application to a plasma-assisted deposition process with run-to-run drifts.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Run- indexed time-varying Bayesian optimization with positional encoding for auto-tuning of controllers: Application to a plasma-assisted deposition process with run-to-run drifts

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.183329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.183329Z digest=sha256:377428bb735c91a484d0d1e122f754818a044f5ceda868d6f444143973b5b5d3

Observation 0abc7c5a-9e60-4a74-9f52-ac1856ce4cc3 · outbound

This paper cites Choksi and Joel A.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Choksi and Joel A

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.303134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.303134Z digest=sha256:33af19b578b271ac11fe7d173d0cce0fd5277a605ee3f7eaf2523297ddf373b2

Observation 0cf21be5-ca6d-415f-a1fa-7091b8de664e · outbound

This paper cites Deep reinforcement learning from human preferences.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Deep reinforcement learning from human preferences

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.442695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.442695Z digest=sha256:4adc31bd35ea1689c2bc79e3183d8562c2ac3ef5f75ac20fd7554da0a911ef6f

Observation a60f030e-1dfb-4b31-9acd-9b4ec375c21c · outbound

This paper cites Model-Based Reinforcement Learning via Meta-Policy Optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.547102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.547102Z digest=sha256:b5db34ff3094c9807087af0a06d986badb4528b88ae1f3dba6d89c933af4cf54

Observation f9a3e710-ab57-484d-820e-0e6b676855d4 · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.650902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.650902Z digest=sha256:fdf3067e7bc1556b8a0526f660fc34ca11c4ed9a10ab2c516d33e51226cee2ea

Observation 7a8d9023-da0c-450b-851c-1a65e2b602fa · outbound

This paper cites Bayesian reinforcement learning in continuous POMDPs with gaussian processes.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bayesian reinforcement learning in continuous POMDPs with gaussian processes

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.783973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.783973Z digest=sha256:bb5e60423190ee0b26abd0580b76173bcef31ef11ef8ccdadb6f818bb3817bc8

Observation 2abae544-0be7-4f26-bd71-66a3cc4029ba · outbound

This paper cites Unexpected improvements to expected improvement for Bayesian optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unexpected improvements to expected improvement for Bayesian optimization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.920530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.920530Z digest=sha256:304145911035e687c4952f5b0da2f018853b72052ea3e43466b0b85f2dbbd1ac

Observation c29a358d-7667-4e93-ad4c-d012972837d5 · outbound

This paper cites Differentiable Expected Hypervolume Improvement for Parallel Multi-Objective Bayesian Optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Differentiable Expected Hypervolume Improvement for Parallel Multi-Objective Bayesian Optimization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:48.997747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:48.997747Z digest=sha256:6383035d8d6293ad891fab09afc1704d767c45397093d1f49d01bc47dcb44d0c

Observation 2c7f02ae-cf66-49e7-80de-7ad56ca92293 · outbound

This paper cites Multi-objective Bayesian optimization over high-dimensional search spaces.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Multi-objective Bayesian optimization over high-dimensional search spaces

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.142742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.142742Z digest=sha256:f24425b01fcf04b489331de161acbdbead3853c9fd4de57fc67c983dc897f783

Observation e9fc999d-c90b-4519-9d40-38384e5579b4 · outbound

This paper cites Osborne, and Eytan Bakshy.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Osborne, and Eytan Bakshy

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.230348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.230348Z digest=sha256:915da089c5976f81ea2330b180d248ef0def985d5f2f0f880c3b90b6c124a060

Observation 0210de10-1e39-41e0-9adb-1464d20d7cd5 · outbound

This paper cites Mixed-Variable Bayesian Optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Mixed-Variable Bayesian Optimization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.336397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.336397Z digest=sha256:208b137f8570aea1103af52d53775e6d8ec838efbca00fd0399f2af63f606bac

Observation c4d791cb-6b20-44d3-85bf-1894a20a4ff3 · outbound

This paper cites Gymnasium robotics, 2024.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Gymnasium robotics, 2024

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.483170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.483170Z digest=sha256:d8e5240fe298dc47cc92eed3fa06ded1e1843b9c277c02f39a547c956de22720

Observation aaa953c0-6213-4f62-bc61-5b55584bdb5f · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.608731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.608731Z digest=sha256:835d4b44853de1e77ad68cf91a788057865a728ea35dd50c28c917f455480c0e

Observation 5cacf78e-d92e-4915-90aa-72c9e3ddf18c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.753278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.753278Z digest=sha256:4ff9b2e277cfa60445ff0cdcd11694b96e0612700abbda9d31db3875532c65c5

Observation e693109d-d6be-4568-a411-8a93d5e9af38 · outbound

This paper cites Dontchev and R.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Dontchev and R

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.848569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.848569Z digest=sha256:78592ad3bd10261fdf08b82b83d14d6d5a212b2d4c24981f1e17856bd85d431e

Observation db3c00b8-6091-4279-a1ed-e52b829f7813 · outbound

This paper cites Additive Gaussian Processes.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Additive Gaussian Processes

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:49.980220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:49.980220Z digest=sha256:5fd1ffea776770039abda10b10ed4293768806258cb887895de56e8f1187318c

Observation d6aaf8e2-8427-4164-8dc6-f5e214697488 · outbound

This paper cites Infinite-Horizon Differentiable Model Predictive Control.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Infinite-Horizon Differentiable Model Predictive Control

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:29:08.505055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:28:50.126635Z digest=sha256:56ab3d4e197900c4b5f765026b44d276883d6f159b92cfa899b77c40af7a784a

Observation 06272063-febb-48c6-81fb-68f7c30fa691 · outbound

This paper cites High-dimensional Bayesian optimization with sparse axis-aligned subspaces.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents High-dimensional Bayesian optimization with sparse axis-aligned subspaces

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.200406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.200406Z digest=sha256:f93c89e6f17ea8fe10bb41f751973e6b06a849c4d340206567a0119b39f1f96c

Observation b2a58269-5ccb-48f9-a0c1-d8760b45e2b1 · outbound

This paper cites Scalable global optimization via local Bayesian optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Scalable global optimization via local Bayesian optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.286025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.286025Z digest=sha256:a83de9089b51b6327ae31f8f2a4a3f69198ebf54291327aabdef8bbdc6eb21a9

Observation 9a96bcb0-c537-4c9e-8eb3-bf71e8c28156 · outbound

This paper cites Policy Gradient Reinforcement Learning for Uncertain Polytopic LPV Systems based on MHE-MPC.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Policy Gradient Reinforcement Learning for Uncertain Polytopic LPV Systems based on MHE-MPC

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.355960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.355960Z digest=sha256:556ab0315b686a4a5cf244f04f112145202f06d326c928d0a256918baac88857

Observation 7f93373b-1917-41bc-9c87-31d31c96d911 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Global convergence of policy gradient methods for the linear quadratic regulator

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.509504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.509504Z digest=sha256:338abdb7ff4099be7bd3d786f308afea5b93e09eb56c6aca1775a44d0f7877da

Observation 77ebcedc-5837-4f2d-8170-dc463a0941f3 · outbound

This paper cites Dynamic Regret of Policy Optimization in Non- Stationary Environments.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Dynamic Regret of Policy Optimization in Non- Stationary Environments

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.622366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.622366Z digest=sha256:3a6ed47dc3e13e131a1a56dfdbf02dc56ae017c8a3e6608e28142cce5a22f149

Observation 8b0f7a8c-ec51-4041-8806-b5bb94663ed9 · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.780885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.780885Z digest=sha256:683ad5a66f762f1bb25defa46f77dd0af34cc455376fce4899a7844cf1f44acd

Observation fd8b8bc1-56b0-40bb-8b46-8c1b77b3b521 · outbound

This paper cites A Tutorial on Bayesian Optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents A Tutorial on Bayesian Optimization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:50.893776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:50.893776Z digest=sha256:9dc861bce9b816f074ca70f55084b38819f1fbb49bcfa33d0c235e1d70b146a4

Observation 773f0b59-9a33-492f-87e0-917ef673b319 · outbound

This paper cites Fr¨ ohlich, Melanie N.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Fr¨ ohlich, Melanie N

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.034133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.034133Z digest=sha256:381b09e479bb7524379a67aeb53f4a839069d33b2b78f05609e37758af674f15

Observation 00e4ea5d-0466-47aa-84ee-a3aa545643c7 · outbound

This paper cites Fr¨ ohlich, Edgar D.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Fr¨ ohlich, Edgar D

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.160277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.160277Z digest=sha256:c3427b2d97afdc3789b234c521a2efeca439e6bce63630fd5e0ea51810d4a101

Observation ff23afdf-4102-4867-b726-540d9165c9c6 · outbound

This paper cites Learning robust rewards with adversarial inverse reinforcement learning,.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Learning robust rewards with adversarial inverse reinforcement learning,

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.312165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.312165Z digest=sha256:d5876333b4ba8a8e9e8cd0963254656c0b7b0114b235a7f34cfe260b041cded3

Observation b831fc45-6678-42d8-8e94-55b9ac6dffd5 · outbound

This paper cites Bayesian Optimization with Inequality Constraints.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bayesian Optimization with Inequality Constraints

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.482326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.482326Z digest=sha256:cc0b06a98f574862572064005b1d25cceb23451d1e325626431721894a194f9f

Observation c193070c-d987-433f-aa08-cbd41e63d02f · outbound

This paper cites Osborne, and Philipp Hennig.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Osborne, and Philipp Hennig

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.668203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.668203Z digest=sha256:60ad8c1175c984d6b333419fa269bf641f6749262c2b8e4451862c941ab505f0

Observation 6f48b991-45db-4c75-862f-f415a93ca696 · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.761659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.761659Z digest=sha256:30e893d8b6178d7aa864ba5d588efb08de157136e39a83814bf6908aab81a2ac

Observation c0a25d99-8986-4e75-bd2a-84f1d70266b4 · outbound

This paper cites Identification for control: From the early achievements to the revival of experiment design.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Identification for control: From the early achievements to the revival of experiment design

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:51.901991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:51.901991Z digest=sha256:20dadc90e22ad230a01da7d0f38054a1357741de0e24ca666599d0f05f408d3f

Observation 378c0845-6a81-4a2d-b0be-d0c77d64fd7a · outbound

This paper cites Multi-objective optimization of a path-following MPC for vehicle guidance: A Bayesian optimization approach.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Multi-objective optimization of a path-following MPC for vehicle guidance: A Bayesian optimization approach

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.092088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.092088Z digest=sha256:4b6a953b08485f28ce6f35be5a14766e00e5b4370690346f0da1008255a77e17

Observation a3714ab1-e776-4a58-9962-2c181298f6d7 · outbound

This paper cites Lawrence.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Lawrence

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.188337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.188337Z digest=sha256:381987fec685ba8748241aeb949f79c4222b0669e9f89f2311a3df296c5337a7

Observation a2018f0b-1014-4e8e-8fb3-56adb8955b9c · outbound

This paper cites Variance Reduction Techniques for Gradient Estimates in Reinforcement Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Variance Reduction Techniques for Gradient Estimates in Reinforcement Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.287056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.287056Z digest=sha256:d3b01e299025c7085cf20dc01bb8d06c96ed179169ab1245382d563e2ff7e332

Observation 18675a04-9d78-45a4-867d-9e776506cec3 · outbound

This paper cites A survey of actor-critic reinforcement learning: Standard and natural policy gradients.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents A survey of actor-critic reinforcement learning: Standard and natural policy gradients

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.429909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.429909Z digest=sha256:a35939135630a4533ac2a29f4a8373350de967f0b8624c7eb54e3067ec42695a

Observation a4aa8710-db08-4f3f-b9b7-30e220691a21 · outbound

This paper cites Learning for MPC with stability & safety guarantees.Automatica, 146:110598, 2022.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Learning for MPC with stability & safety guarantees.Automatica, 146:110598, 2022

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.587647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.587647Z digest=sha256:60c932bc2e91bdae37a50e680e4a9a7c6b72cbe936a8226224339c17343e4f4e

Observation cfa3db73-79b8-4c35-9657-0102213c6f3f · outbound

This paper cites Safe Reinforcement Learning via Projection on a Safe Set: How to Achieve Optimality? IF AC-PapersOnLine, 53(2):8076– 8081, 2020.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Safe Reinforcement Learning via Projection on a Safe Set: How to Achieve Optimality? IF AC-PapersOnLine, 53(2):8076– 8081, 2020

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.681105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.681105Z digest=sha256:f51ede2c6152b35c10a5d24502eeff36d9cc69087a12a6b6158e2d98448162b9

Observation 59cf1326-66f8-46b6-a59b-94dda879d8e5 · outbound

This paper cites Data-Driven Economic NMPC Using Reinforcement Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Data-Driven Economic NMPC Using Reinforcement Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.801739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.801739Z digest=sha256:de4c0e8c421a11e49e3597a85f5d178f11baed9caeec480f6fedc327be281024

Observation 42ec66f3-8de9-44d8-9d9f-8b8044d6e82d · outbound

This paper cites Reinforcement learning for mixed-integer problems based on MPC.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Reinforcement learning for mixed-integer problems based on MPC

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:52.917251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:52.917251Z digest=sha256:082230ff2b10a9962a51d4c56a64840ed3a98391c4c0fdea760f0da532664c1a

Observation 4272abc0-ebae-4cb2-b051-432f2bef1fff · outbound

This paper cites Reinforcement learning based on MPC and the stochastic policy gradient method.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Reinforcement learning based on MPC and the stochastic policy gradient method

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.032903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.032903Z digest=sha256:ab24d249e063a49baca0fd35d3b5b23328181450b0e86363e909eaedb20cbd81

Observation 1d02cf75-c4a2-4bbf-b276-f7fccaa57a8d · outbound

This paper cites Guerreiro, Carlos M.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Guerreiro, Carlos M

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.215348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.215348Z digest=sha256:a83a9aa948b17f82ae23b706499ffe73b6f3b5075b3a8e1f6b478c83a0bf8762

Observation bf4d0789-a6ac-4db9-9087-4b7688807207 · outbound

This paper cites Evolutionary optimization of high- dimensional multiobjective and many-objective expensive problems assisted by a dropout neural network.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Evolutionary optimization of high- dimensional multiobjective and many-objective expensive problems assisted by a dropout neural network

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.364738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.364738Z digest=sha256:3f0ab91c0c7c544fbe9e3719a4867f60315ec5b7a2d66d58370437803570de9b

Observation 5d57e077-d5c2-48f0-b871-9801f5aabbec · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.451248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.451248Z digest=sha256:2a7810eb72a9bee9ae06d5b23ac36314fae72e5f5a8642040b698705fa023503

Observation 8f6b6404-08fd-4314-ac2b-a056f07a74a9 · outbound

This paper cites Mastering diverse control tasks through world models.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Mastering diverse control tasks through world models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.616540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.616540Z digest=sha256:57517ab34f9dd625d35f4293ab08ff6a0b9b6874ec772eab8a7ae508ad7208ab

Observation 985e64d1-e6f0-46ea-96ad-c72658422f6a · outbound

This paper cites M¨ uller, and Petros Koumoutsakos.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents M¨ uller, and Petros Koumoutsakos

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.746417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.746417Z digest=sha256:888f9c0a3bf6046c76c58be41aed32dac64e8d3560ba684994b3aa2ed0e76e50

Observation 6a93eed2-b3e2-4e98-88fa-2808674b1348 · outbound

This paper cites Hayes et al.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Hayes et al

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:53.896075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:53.896075Z digest=sha256:ea6292ab2bed825f080c1ecb923edaec6a030e9426c861c24986bf20eef73c56

Observation 424688ee-7bfe-4e01-87fd-262b4592691f · outbound

This paper cites Deep Gaussian process for multi-objective Bayesian optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Deep Gaussian process for multi-objective Bayesian optimization

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.067409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.067409Z digest=sha256:5bb4cde342749b7b102e819384a5502d82da61b4348b0c1a30797ebc84648a9d

Observation 13cb9411-1ac4-41b8-b7b5-1e2b911a13c2 · outbound

This paper cites Wabersich, Marcel Menner, and Melanie N.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Wabersich, Marcel Menner, and Melanie N

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.160670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.160670Z digest=sha256:42bd91886f27e2c89d9edd0678fb043b012bddde19ea38b879348076226dcf6b

Observation a4e564d1-66e1-4475-89a6-404301392b3b · outbound

This paper cites Stability-informed Bayesian Optimization for MPC Cost Function Learning.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Stability-informed Bayesian Optimization for MPC Cost Function Learning

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.218919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.218919Z digest=sha256:8bba344df701347480917f8c9a69177d3b782a257714513e6106f4cddeaabba8

Observation 5c7c6898-a4e8-4634-ad24-e03b9aee322f · outbound

This paper cites Multi-objective Bayesian optimisation over sparse subspaces for model predictive control of wind farms.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Multi-objective Bayesian optimisation over sparse subspaces for model predictive control of wind farms

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.267010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.267010Z digest=sha256:0cca8b534e671841a8be5de56a6f8aa96b7d3d0f41519f736030f6ab070b6365

Observation 02b17292-dd46-4f76-9494-e62e8d7762b2 · outbound

This paper cites Multilayer feedforward networks are universal approximators.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Multilayer feedforward networks are universal approximators

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.323315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.323315Z digest=sha256:163f20aecc873d451a35bbb47cad753bf1f9100de9cf763999a492ed169d14e8

Observation 8b35d3e4-8d5b-418d-9257-caf1f423dee2 · outbound

This paper cites Reinforced Few-Shot Acquisition Function Learning for Bayesian Optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Reinforced Few-Shot Acquisition Function Learning for Bayesian Optimization

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.417150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.417150Z digest=sha256:073fab40b02fd3cec6768280e5ce5517fc6f41a98762a7b6261a03bc26d4de09

Observation e5df7bd6-1aa3-4c4c-9b3f-a38304e3b418 · outbound

This paper cites Toward a theoretical foundation of policy optimization for learning control policies.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Toward a theoretical foundation of policy optimization for learning control policies

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.521695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.521695Z digest=sha256:1f0d48dcadcbbc6e29544f0253f4aaf6aa932bd15013bc9933378745d778f5c8

Observation e69dc31d-ede3-4223-93e5-78dfad687a2a · outbound

This paper cites BOFormer: Learning to Solve Multi-Objective Bayesian Optimization via Non-Markovian RL.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents BOFormer: Learning to Solve Multi-Objective Bayesian Optimization via Non-Markovian RL

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.642701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.642701Z digest=sha256:4ea8e129d7ef7798147033b76ea2d7d716c7f3e60c461b2b1f0a505ac33dd2cb

Observation c607f2fd-834d-4132-b10b-08553b2fbb0c · outbound

This paper cites Hoos, and Kevin Leyton- Brown.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Hoos, and Kevin Leyton- Brown

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.744749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.744749Z digest=sha256:7ca0d7642b6e764a4f8d722fb5a5b79576f66eda1d1d6af4122b3ee486d2bf56

Observation 8409ded2-003a-4d9e-8ef3-5ff90c1ca3a3 · outbound

This paper cites J., Santosh Penubothula, Chandramouli Kamanchi, and Shalabh Bhatnagar.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents J., Santosh Penubothula, Chandramouli Kamanchi, and Shalabh Bhatnagar

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.823943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.823943Z digest=sha256:5c427215f3d5b60ab548f44043148a429352f277c361d2c1d5a2f5e4a85ba9b9

Observation 52639965-2c9f-4507-b5a9-096583ae1d7b · outbound

This paper cites When to Trust Your Model: Model-Based Policy Optimization.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents When to Trust Your Model: Model-Based Policy Optimization

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:54.925681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:54.925681Z digest=sha256:0779d117296d2c8d9c9916e5ab6646e70427f9f7116a871a83640cae05bd071e

Observation ad30fe06-42b9-4105-89e7-eb02943d00dd · outbound

This paper cites Bilevel optimization: Convergence analysis and enhanced design.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Bilevel optimization: Convergence analysis and enhanced design

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.013776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.013776Z digest=sha256:731ac3688aec80e0688c96d364551feba403d2cf570917a3d023281c5dde3f53

Observation 65c7e323-be5f-46ed-9849-e3a31a173c21 · outbound

This paper cites BINOCULARS for efficient, nonmyopic sequential experimental design.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents BINOCULARS for efficient, nonmyopic sequential experimental design

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.092899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.092899Z digest=sha256:569600d88716e78262fee4c4422d94e3031289dc6ab2ac6400d588049c9376f1

Observation a8346639-dbf4-4af7-bd39-1ca183a70804 · outbound

This paper cites an unresolved cited work.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Unresolved cited work

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.203017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.203017Z digest=sha256:f47e4cb7011a21707dc41343c5e57b9453872919c9575adb240412ab5c447eb1

Observation d787f3ea-9142-48ca-a59d-9e089b7a6597 · outbound

This paper cites Pontryagin Differentiable Programming: An End- to-End Learning and Control Framework.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Pontryagin Differentiable Programming: An End- to-End Learning and Control Framework

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.303038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.303038Z digest=sha256:9693c616adf0dcaf2cf46c42f9b0ab427355635269e7211d6b632ef7a0f34689

Observation 56f647ea-4e48-4fdd-b1df-7239a29065bf · outbound

This paper cites Data-efficient reinforcement learning with probabilistic model predictive control.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Data-efficient reinforcement learning with probabilistic model predictive control

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.374609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.374609Z digest=sha256:1df6b3e70c8075ad36da6f842918333665ab3a9b50a4b0cf3d7e86397318124d

Observation 00b79c4c-7343-41b9-94ce-66e91a59bdc3 · outbound

This paper cites A review on genetic algorithm: past, present, and future.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents A review on genetic algorithm: past, present, and future

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.476462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.476462Z digest=sha256:0ae93a52b219bcf7e3e9651389703ce52d145c4fd40b476434e2ac7d1c2dd020

Observation 2c16d47a-591c-4a9f-ae6c-590fd5187d12 · outbound

This paper cites Lekkas, and S´ ebastien Gros.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Lekkas, and S´ ebastien Gros

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.557377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.557377Z digest=sha256:5fa72981bc450c66c1800b98178524091425dc239bc2e57b5fc1e42df6581396

Observation 1c661c43-f00b-45f3-a849-90fe67464ae9 · outbound

This paper cites Doyle III.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Doyle III

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.649869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.649869Z digest=sha256:670999832ec7a015097a2b4f09eb37aca2e8243ce892a09c7d14ffedea4124f8

Observation 44ec2c76-757e-4fa2-9cd5-f570cd561f56 · outbound

This paper cites Huynh, Ali Mesbah, and Joel A.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Huynh, Ali Mesbah, and Joel A

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.743538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.743538Z digest=sha256:6ff4283c60f5074e18a1e10812ff73f838e89aa0b7c44dec23d23396a36eacd1

Observation 109af6cc-e0b4-4957-8b30-076b5ab235a9 · outbound

This paper cites Lagoudakis and Ronald Parr.

Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents Lagoudakis and Ronald Parr

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-06T16:28:55.851695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:28:55.851695Z digest=sha256:86bb0ca7da4a1ef461facd131375367c4d46d3da0f084214b5b886a97c7f18c2

Pith citing papers

Observation b10582df-bbf4-4826-8a45-9d474592ab6d · inbound

FlexPath: Adapting Learned Connectivity Guidance to Path Preferences cites this paper.

FlexPath: Adapting Learned Connectivity Guidance to Path Preferences Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:17:30.888870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T16:41:29.124842Z digest=sha256:545679c299c8843d1bde28ffaf9162f60609a257a914a12bac65d688664d3312

Observation 58d072fe-1c40-4f82-8c23-b2706d46233e · inbound

FlexPath: Adapting Learned Connectivity Guidance to Path Preferences cites this paper.

FlexPath: Adapting Learned Connectivity Guidance to Path Preferences Model-free Reinforcement Learning for Model-based Control: Towards Safe, Interpretable and Sample-efficient Agents

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T02:14:58.076642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:14:58.076642Z digest=sha256:15d9c4d77b390490f61c6a47ab00de4420efa07326ee6d7c2363e7e56aae001b