Pith. sign in

Paper Citation Record · LEDGER

Model-based Lookahead Reinforcement Learning

As of 20 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:1908.06012.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.06012 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T13:17:37.396552Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact0
  • verified fuzzy36
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 776294e9-624a-46b1-b34f-dec4ae894125 · outbound

This paper cites P., Bertsekas, D.

Model-based Lookahead Reinforcement Learning P., Bertsekas, D

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.889625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.231442Z digest=sha256:b56f798bd2d6771ff2ffff9747a769b8994734e8996ca135d5faeb4fb063d558

Observation fabf6bcb-aa9a-48b9-bbd0-ee0dd77e7150 · outbound

This paper cites Openai gym, 2016.

Model-based Lookahead Reinforcement Learning Openai gym, 2016

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.878927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.235927Z digest=sha256:23db77ddf3fbfc16234f8f072488b9e726786c6df0a29e5346ac12212f40f08d

Observation 79be8fd6-875e-4cf9-8702-40e5ff0e7197 · outbound

This paper cites Sample-efficient reinforcement learning with stochastic ensemble value expansion.

Model-based Lookahead Reinforcement Learning Sample-efficient reinforcement learning with stochastic ensemble value expansion

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.867393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.239520Z digest=sha256:5d693b1075cfc0c1eefb9b04caf2e96b00c93c40836349c342f8d24bac622cb0

Observation 663cb465-4bec-4723-a978-1b559c5243f4 · outbound

This paper cites an unresolved cited work.

Model-based Lookahead Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:17:37.856698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.243699Z digest=sha256:65cad43dc05a27a497626ed67e14b4fdebdd8602f98465ee848b643da7033c43

Observation 253a75bb-e7f7-4f74-a597-59d921778c19 · outbound

This paper cites Path integral guided policy search.

Model-based Lookahead Reinforcement Learning Path integral guided policy search

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.846467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.248363Z digest=sha256:8b4d8f5bae5576af4e253f8f0fad8dbf2dc4ede614bdc238a6678abfbba9688d

Observation db5b2ffc-cfc0-4def-854a-540962279506 · outbound

This paper cites Deep reinforcement learning in a handful of trials using probabilistic dynamics models.

Model-based Lookahead Reinforcement Learning Deep reinforcement learning in a handful of trials using probabilistic dynamics models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.836699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.252379Z digest=sha256:73c8e416e8cf6e3f34b469ee6d8d5840cbbedb8140c869d429e806ebf38a2b80

Observation baaf0c54-ea46-4b32-8d0b-f287051b7d20 · outbound

This paper cites Model-Based Reinforcement Learning via Meta-Policy Optimization.

Model-based Lookahead Reinforcement Learning Model-Based Reinforcement Learning via Meta-Policy Optimization

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T13:17:37.256111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:17:37.256111Z digest=sha256:b29a4d7866b6581846c9ede4dd63138e9c2c56edba8a118aa47626c49414d36d

Observation 419700fa-9bdc-4041-8247-5369737e0ac6 · outbound

This paper cites and Rasmussen, C.

Model-based Lookahead Reinforcement Learning and Rasmussen, C

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.826065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.259734Z digest=sha256:1ad03c7211b52ac38df7089e4c1c700cbd99d8d50cc208054ef2048fae2142d4

Observation 56bfb4ff-688c-4536-908c-472cd95ce25e · outbound

This paper cites S., Landau, S., Leese, M., and Stahl, D.

Model-based Lookahead Reinforcement Learning S., Landau, S., Leese, M., and Stahl, D

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.816275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.264166Z digest=sha256:7897df99f0bf44fdabec3c18bf7305111b8b502a20fce02441f8c5458ebd0cda

Observation f18e2050-0ec2-43d4-b79d-1a4833bbc2ea · outbound

This paper cites Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning.

Model-based Lookahead Reinforcement Learning Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T13:17:37.267563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:17:37.267563Z digest=sha256:667be981c373b645821d2fc2de3c28bec12a77bf13855018a69e1014bd5247fd

Observation 17609a03-51e6-4ea6-ac2e-5624f0034d74 · outbound

This paper cites E., Prett, D.

Model-based Lookahead Reinforcement Learning E., Prett, D

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.807117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.271862Z digest=sha256:79b8a506da7d88f44883024625d126ad5463c7bcc2ebf6e6073e43538200f810

Observation 8b727540-6c4d-4ceb-947c-3d31e4fce761 · outbound

This paper cites Continuous deep q-learning with model-based acceleration.

Model-based Lookahead Reinforcement Learning Continuous deep q-learning with model-based acceleration

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.796375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.275335Z digest=sha256:f69795614436c441873138077e052497ff30f976aa118d134ee6b5c13b9d6584

Observation 0d1db2bb-987b-4666-b769-6a848141db83 · outbound

This paper cites and Boedecker, J.

Model-based Lookahead Reinforcement Learning and Boedecker, J

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.785524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.279350Z digest=sha256:6c15bf71f616bce5dde8388600a7c5e4b0bab0fb28db93b065d7a3e58470aba6

Observation 62f1c7ac-bbd8-4ffc-9363-3849a644eb33 · outbound

This paper cites and Deisenroth, M.

Model-based Lookahead Reinforcement Learning and Deisenroth, M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.775589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.282549Z digest=sha256:ba71ade77bcc523680fc902b2ab180380ea4dd7bb1d5ae051214823686b8fdc9

Observation 5358cb5e-add2-4517-92d0-88e34577ed48 · outbound

This paper cites an unresolved cited work.

Model-based Lookahead Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:17:37.765337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.286557Z digest=sha256:ba82be021d0c8d646a166b24b53df5ee0c8c92c886f1a50deae24a38d055cf8f

Observation 5616f29b-bf39-4dec-a9f4-6b963d9b0ceb · outbound

This paper cites Model-ensemble trust-region policy optimization.

Model-based Lookahead Reinforcement Learning Model-ensemble trust-region policy optimization

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.755074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.289984Z digest=sha256:0afdc9aefc76f474fb14d82b2edce0926b4237d93a4884920deaf51a9d339717

Observation a455a082-acb5-44ae-9899-2397c76d6ff5 · outbound

This paper cites and Abbeel, P.

Model-based Lookahead Reinforcement Learning and Abbeel, P

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.744962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.294269Z digest=sha256:24f741193f02427a72020d38736edcfb833f4afadf6d715281eebe7001bc70f1

Observation 8b8bcc0f-1d01-4f8b-b510-07c1bf1bd264 · outbound

This paper cites and Koltun, V.

Model-based Lookahead Reinforcement Learning and Koltun, V

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.734923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.298373Z digest=sha256:c34455d8872fa926c6c42a692b65c3caabc3afdf7e4f52d392a7b6b8a470e0b3

Observation 75a62b1e-2405-4a33-b68e-ee12129cb0ee · outbound

This paper cites P., Hunt, J.

Model-based Lookahead Reinforcement Learning P., Hunt, J

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.725309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.301710Z digest=sha256:3a81a1ca5464f39c996ff34790a4a5cca3045ed228701f37b90f7b52972b6781

Observation fbae266b-1913-4a45-9076-ca6ab5126c0d · outbound

This paper cites Plan online, learn offline: Efficient learning and exploration via model-based control.

Model-based Lookahead Reinforcement Learning Plan online, learn offline: Efficient learning and exploration via model-based control

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.715706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.305679Z digest=sha256:c08e62578bbea2f54e3f482c953758d596d224353d6f19232938b9c6886fac65

Observation 73cf11d9-66b6-44a8-9c0c-bdf1439f05cd · outbound

This paper cites Human-level control through deep reinforcement learning.

Model-based Lookahead Reinforcement Learning Human-level control through deep reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.705794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.309048Z digest=sha256:4c95cdb86529e0898bffc52b7fafcdefe69ea5b0804598f049c92099a947620b

Observation a1865dd4-34dd-4762-b6dd-ee2de984ae6d · outbound

This paper cites Asynchronous methods for deep reinforcement learning.

Model-based Lookahead Reinforcement Learning Asynchronous methods for deep reinforcement learning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.695918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.312694Z digest=sha256:9d469b0cce34eb062e7d54d79a4d881077b149ccc6817764fc2f106c62ef4c14

Observation c68fccc3-3412-44ce-b418-421eb2ef496f · outbound

This paper cites S., and Levine, S.

Model-based Lookahead Reinforcement Learning S., and Levine, S

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.684445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.316047Z digest=sha256:9ebf0755572a180f8447ec08f673b76072d1c36546d19d5bae9b8bc37fffb0d6

Observation d4d28720-0592-4490-a536-5e382e40fa1d · outbound

This paper cites Value prediction network.

Model-based Lookahead Reinforcement Learning Value prediction network

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.674084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.320395Z digest=sha256:8cb1b7410c4dce9cc5bdb13c2b1379d5c240c22930ce1b4893136971f1b92b01

Observation ac5f6293-98cb-4644-be0e-142b86ac56b5 · outbound

This paper cites Temporal difference models: Model-free deep RL for model-based control.

Model-based Lookahead Reinforcement Learning Temporal difference models: Model-free deep RL for model-based control

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.662538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.324112Z digest=sha256:c6143b128b12b46cd0de8c48dd3edee6b0da853d83db4bd92559511a60e91392

Observation f1718961-02c7-49f5-8569-33c92c5bf7aa · outbound

This paper cites Imagination-augmented agents for deep reinforcement learning.

Model-based Lookahead Reinforcement Learning Imagination-augmented agents for deep reinforcement learning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.651777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.327996Z digest=sha256:da898752d8aafa7f4a90cb411f058ee3315ece336a2688e575552dfe9020e76d

Observation 39317062-92d7-4dfb-94ff-e2981f9a1b85 · outbound

This paper cites an unresolved cited work.

Model-based Lookahead Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:17:37.641438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.332256Z digest=sha256:befc78d89e8172d21ed654f8f38c513539511452ee67ba3dd7bfee18692df068

Observation a9447858-dad8-455c-840a-2e3088c36866 · outbound

This paper cites The cross-entropy method for combinatorial and continuous optimization.

Model-based Lookahead Reinforcement Learning The cross-entropy method for combinatorial and continuous optimization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.631844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.336076Z digest=sha256:bc43a636d3160406ce62df3e5bdf5c89c27169ae2726578efeaa07754345bd72

Observation 206744e1-bfc8-4557-a0f1-b9801517f350 · outbound

This paper cites Trust region policy optimization.

Model-based Lookahead Reinforcement Learning Trust region policy optimization

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.621638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.339587Z digest=sha256:19f13a7d0b1e90929c5f8752270b8681547d82e7142577ca9b4cd0a81252f712

Observation bb6bb807-7553-4ba6-968b-46024598ae29 · outbound

This paper cites High-dimensional continuous control using generalized advantage estimation.

Model-based Lookahead Reinforcement Learning High-dimensional continuous control using generalized advantage estimation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.610305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.342974Z digest=sha256:3b23a61562ed16dc7b95e45fff2a648773be7d60b5a7766765e6fd5a6e576748

Observation 5b4a165f-b097-4e28-a023-39fa80f1ef2a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Model-based Lookahead Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T13:17:37.346297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:17:37.346297Z digest=sha256:481f84c915dc323e6fd0962c9c8baf2f28f1d076c0397638fc6ee89f13f710cc

Observation 1daab8d8-610c-45b8-87a9-8c9881af384a · outbound

This paper cites S., and Müller, M.

Model-based Lookahead Reinforcement Learning S., and Müller, M

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.599830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.350120Z digest=sha256:c91520614f0a954745d235b0a15a171691f0fd12247248b9625bf5ba32ead0a4

Observation 04afc63b-720a-4d19-be61-ed2d927076be · outbound

This paper cites J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V ., Lanctot, M., et al.

Model-based Lookahead Reinforcement Learning J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V ., Lanctot, M., et al

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.590146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.353658Z digest=sha256:ef1f650c25b9b0b7d2d21a6046502bfbaef5e4f6928e00ad165f0b71cc6f124e

Observation e847c61d-8b81-484e-982a-f3fda0f64d4d · outbound

This paper cites an unresolved cited work.

Model-based Lookahead Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:17:37.580310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.357524Z digest=sha256:e689d8e10b658ece8056dabf5ce080f38a1bdf4ee13b890cf88a49ad90e21fad

Observation 38e879ad-ee36-4b66-8541-fa97091598cc · outbound

This paper cites an unresolved cited work.

Model-based Lookahead Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:17:37.570878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.361972Z digest=sha256:2faf9e52fe8cd50115b65f05187b81cc713af7b8ce7e5b54686ec2c9776a8d26

Observation 8ce5b5d7-cbcd-4d57-92c7-81690b31f1fa · outbound

This paper cites S., McAllester, D.

Model-based Lookahead Reinforcement Learning S., McAllester, D

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.560709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.365475Z digest=sha256:0f7e63eb2ce178cab440d86fba8c67f07095d29c159f2628b6fe15e7eff2c016

Observation d841f841-efcc-425f-a60a-f814578cc467 · outbound

This paper cites S., Szepesvári, C., Geramifard, A., and Bowling, M.

Model-based Lookahead Reinforcement Learning S., Szepesvári, C., Geramifard, A., and Bowling, M

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.550369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.368753Z digest=sha256:ed5b70b4eec904076934d3ffa9cf22f709392509d1d3a9a81ccafafe7cdefa27

Observation 48f487be-f315-421f-89be-0fea53bf9a3f · outbound

This paper cites Value iteration networks.

Model-based Lookahead Reinforcement Learning Value iteration networks

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.539359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.372152Z digest=sha256:da2415ddcc0cf58c7501db71b29f858fc011b28843fec4023badec1c07714212

Observation af888445-fc10-4297-ba8d-c89d6c3db530 · outbound

This paper cites Synthesis and stabilization of complex behaviors through online trajectory optimization.

Model-based Lookahead Reinforcement Learning Synthesis and stabilization of complex behaviors through online trajectory optimization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.527957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.376395Z digest=sha256:22bb1eb49b3bea980915afb2158ef75b26adee3a87297f76b5a5e44a287c1cff

Observation de43765c-0643-4f15-9614-49ef7830ea3d · outbound

This paper cites Control-limited differential dynamic programming.

Model-based Lookahead Reinforcement Learning Control-limited differential dynamic programming

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.515660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.379869Z digest=sha256:9c2d8550d179d48065da39041ef7af28dee73b6bb14eeeacc8b6002430f5e1dd

Observation 3eadc33b-c3f6-455d-9ec8-dd2eba3fa775 · outbound

This paper cites and Li, W.

Model-based Lookahead Reinforcement Learning and Li, W

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.502924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.383098Z digest=sha256:ccd159efff6104695373a84145e57d5f889546bcf681c141cd22704d1afca9a7

Observation 53a07cd6-27fe-4312-952c-055db9881305 · outbound

This paper cites Mujoco: A physics engine for model-based control.

Model-based Lookahead Reinforcement Learning Mujoco: A physics engine for model-based control

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.492279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.386536Z digest=sha256:0f5c24f1c961add91c03f76e7627c9b1b0c614ae6b093310a4d1d5c20baa5d11

Observation 353316e4-5ba0-4ae4-9ea6-1e772cb52d01 · outbound

This paper cites M., Boots, B., and Theodorou, E.

Model-based Lookahead Reinforcement Learning M., Boots, B., and Theodorou, E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.480079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.389840Z digest=sha256:be35a74ab4c1433395842a93f7b6d33f89304366844a92099c6e5345bde54993

Observation 5759920e-d39d-4673-8b30-067003c84a6f · outbound

This paper cites an unresolved cited work.

Model-based Lookahead Reinforcement Learning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-14T13:17:37.468351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.393125Z digest=sha256:fbdcdbaf4ef60656703410b79c34c0e6fbedf239eee3f2aa7e8e4b1c29d7855e

Observation c0b22320-15d7-462f-8600-aa8f9b155c09 · outbound

This paper cites Learning deep control policies for autonomous aerial vehicles with mpc-guided policy search.

Model-based Lookahead Reinforcement Learning Learning deep control policies for autonomous aerial vehicles with mpc-guided policy search

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:17:37.457788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-14T13:17:37.396552Z digest=sha256:7c1f592effe55731a9c91ab117c376b5d3313925441f1c9d55be469770939b3e

Pith citing papers

No inbound Pith citation observations are available.