Pith. sign in

Paper Citation Record · LEDGER

Temporal Difference Learning for Model Predictive Control

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2203.04955.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2203.04955 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:38:30.271605Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:00:06.256105Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b86049ce-e128-480b-999e-31cd066c7fcb · inbound

DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning cites this paper.

DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning Temporal Difference Learning for Model Predictive Control

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T16:06:09.742215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T16:06:09.448517Z digest=sha256:e250d93254f7230d2ab32c366072e2e2fe1c6234952e555640f5a63ad6223788

Observation 4249080f-d4b3-4d86-a741-110d2eac392b · inbound

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint cites this paper.

TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T04:38:30.271605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:38:30.271605Z digest=sha256:1066c8837c179959d66e426d8d35dba0c2d909c7e00118c766e0cd4160201aec

Observation eab4e933-d8a5-4a34-a61d-f518ae8a57d6 · inbound

WoMAP: World Models For Embodied Open-Vocabulary Object Localization cites this paper.

WoMAP: World Models For Embodied Open-Vocabulary Object Localization Temporal Difference Learning for Model Predictive Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:58.397406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:58.397406Z digest=sha256:3d5f11fd040f8eba5c4d43e3d0e5f6bb2956a4fa9fedef84bffb58c80b8c7cf0

Observation a750b801-2ecd-4803-8de8-21424d59e889 · inbound

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning cites this paper.

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:52:15.261991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:48:28.980868Z digest=sha256:900d0bd518f51bc5f8424982b94a7199af6491e3f5e0646366ff156003bcf997

Observation ae0cd45e-f4f7-48f4-b121-c83b93de9fc5 · inbound

Real-Time Execution of Action Chunking Flow Policies cites this paper.

Real-Time Execution of Action Chunking Flow Policies Temporal Difference Learning for Model Predictive Control

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:18:51.670797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T14:18:51.613045Z digest=sha256:3e246f5190f3d4b8b1d21580864b96dd9217a48e67b00f4a98e60be01083458a

Observation 5754e964-70a6-4abd-a996-3b33939c6594 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Temporal Difference Learning for Model Predictive Control

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:33:50.866977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:c5c71f06b3933879366e2f11ed1976557cbd4b4513b933788080e6346fa7b017

Observation 1a33e769-05e3-4b25-9260-985923d1599b · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Temporal Difference Learning for Model Predictive Control

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:17:14.161907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:c07db63efe54c3dbf658f6fcb455424f156ec3b86f4d836029cbf6c477b06f0e

Observation 2499a6b0-ac18-4471-ab79-133415da982d · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond Temporal Difference Learning for Model Predictive Control

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:22.074261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:22.074261Z digest=sha256:877d62a04af71f05d56a87fbc2f3bc07228f77cfdf1ae73d43195c31ce554b87

Observation ffd82329-753f-467b-93d4-035e47ca5a57 · inbound

Investigating Lagrangian Neural Networks for Infinite Horizon Planning in Quadrupedal Locomotion cites this paper.

Investigating Lagrangian Neural Networks for Infinite Horizon Planning in Quadrupedal Locomotion Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:28.994599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:28.994599Z digest=sha256:20ee135c3f1961c2938503951f2b42c95ea30e27a29eec57649baa5efd23c4b9

Observation 758f57f7-b5a3-47c9-8015-bdaf72ee23d4 · inbound

M3PO: Massively Multi-Task Model-Based Policy Optimization cites this paper.

M3PO: Massively Multi-Task Model-Based Policy Optimization Temporal Difference Learning for Model Predictive Control

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:48.993986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:48.993986Z digest=sha256:ad624f99809676cb61a098854a02cf7b48c77c35ee7552de51f37aca9d457658

Observation b20647d5-91c5-4072-a4c9-6280a9e409cc · inbound

Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease cites this paper.

Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease Temporal Difference Learning for Model Predictive Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:12:58.596909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:12:58.596909Z digest=sha256:465e9637aabe1dc518b900d01850df55279deb72b75be187f6535a63b1de8dac

Observation 6539a871-cc50-41c3-9916-4241e83ac012 · inbound

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution cites this paper.

Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution Temporal Difference Learning for Model Predictive Control

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:06:33.769454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:06:33.769454Z digest=sha256:89a13fb1243164d3fb2433bfd6575e6bb82e662b5581dc90d7a16f9e077f7623

Observation a0943cec-78f7-44e1-958e-ca6abdf9043f · inbound

Arnold: a generalist muscle transformer policy cites this paper.

Arnold: a generalist muscle transformer policy Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T16:41:53.486675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:41:53.486675Z digest=sha256:fcbae5ad770126136006f075c4ba0f0598538d60d2c93e90847008a24dbb7924

Observation 7b387953-eda6-4620-9c97-6611d230f78c · inbound

High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics cites this paper.

High-Precision and High-Efficiency Trajectory Tracking for Excavators Based on Closed-Loop Dynamics Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:31:33.478521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T15:30:19.253933Z digest=sha256:9b309214c862a4db528f72360d2ef0420a5171190236fecb0bca47ab6ffd48c3

Observation b0bb3a76-b285-49d5-9221-c98dfe250bd5 · inbound

Model-Based Reinforcement Learning under Random Observation Delays cites this paper.

Model-Based Reinforcement Learning under Random Observation Delays Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:36:28.673612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T14:35:04.201252Z digest=sha256:91c921e4cd95b6005d6799f9e2f941ce37499694be055d5eafe054232be28fcc

Observation 02524fc5-bb06-4565-9c41-dd5ae15b49d1 · inbound

D2 Actor Critic: Diffusion Actor Meets Distributional Critic cites this paper.

D2 Actor Critic: Diffusion Actor Meets Distributional Critic Temporal Difference Learning for Model Predictive Control

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:35:29.447709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T07:31:27.330951Z digest=sha256:4d548eb65d859564d2df4c78260b0b31b5758fd057b104d1aabd37623d258ef4

Observation 1e85fd0c-eeb7-4b03-b1b0-084afa60a574 · inbound

Ctrl-World: A Controllable Generative World Model for Robot Manipulation cites this paper.

Ctrl-World: A Controllable Generative World Model for Robot Manipulation Temporal Difference Learning for Model Predictive Control

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T01:14:10.399192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T01:14:10.174044Z digest=sha256:643a9e19d7cf61faf97b69912468f26198125ebee1bae97423f4097a21a99a8b

Observation e07f23f2-a065-4dc2-a677-93eb3386fc5f · inbound

Next-Latent Prediction Transformers Learn Compact World Models cites this paper.

Next-Latent Prediction Transformers Learn Compact World Models Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T07:25:29.528570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T07:21:04.684347Z digest=sha256:a3a93d60c525648984fa800df90c1398c8a7568c51e8c4ae5d6b92510aedf6b6

Observation 5557c7da-06d9-48d5-97af-6e146eac1cb0 · inbound

Next-Latent Prediction Transformers Learn Compact World Models cites this paper.

Next-Latent Prediction Transformers Learn Compact World Models Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T23:27:55.490000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:27:55.490000Z digest=sha256:3684bc34796d886c9bddff2bacfeeec52f3e68de94cdc9d0c7c556fa5f7416ee

Observation 5ebb0c64-7a51-49cc-8dd9-3cc23dc5e40b · inbound

CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance cites this paper.

CoRL-MPPI: Enhancing MPPI With Learnable Behaviours For Efficient And Provably-Safe Multi-Robot Collision Avoidance Temporal Difference Learning for Model Predictive Control

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T22:45:05.283619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:45:05.283619Z digest=sha256:28511b016bbbb20dfd11bb10fa29cf369728a720eac3bfcdbc914db8db2514e9

Observation bc98658b-637c-4452-9ece-d4ffce8d7d84 · inbound

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning cites this paper.

Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning Temporal Difference Learning for Model Predictive Control

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:50:12.913336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T14:50:12.804707Z digest=sha256:1aef253e40f9ac5fc6f59a2db9e999eb8ae748607547af376f6d837652613898

Observation eafb0b50-f16f-42ca-8105-9781c824625a · inbound

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning cites this paper.

Beyond Test-Time Memory: State-Space Optimal Control for LLM Reasoning Temporal Difference Learning for Model Predictive Control

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-15T12:09:16.468055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:09:16.468055Z digest=sha256:2eb34df1a37a8cbb8cb6c95528c51ecb41f880406539fba0d9ce925b61c31a07

Observation e8b490cc-ed7f-414a-ba21-c9bfa54c004e · inbound

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC cites this paper.

RAY-TOLD: Ray-Based Latent Dynamics for Dense Dynamic Obstacle Avoidance with TDMPC Temporal Difference Learning for Model Predictive Control

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.799296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T08:43:35.563513Z digest=sha256:628154f1f44b329c1f385a803b92b03bd4b0819b63afb40ff7fd54b1c3cf655d

Observation f352925b-cc14-4e46-81e1-416b17436e9b · inbound

TRAP: Tail-aware Ranking Attack for World-Model Planning cites this paper.

TRAP: Tail-aware Ranking Attack for World-Model Planning Temporal Difference Learning for Model Predictive Control

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:56:04.901730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:15:45.407680Z digest=sha256:4e45f9288619a73bb0b549b2c3871a6c7e6fa12f4697376043492553fe9d95a8

Observation 4ff10b9a-d7d1-40fe-975d-2e3dd7cc63b3 · inbound

Learning Visual Feature-Based World Models via Residual Latent Action cites this paper.

Learning Visual Feature-Based World Models via Residual Latent Action Temporal Difference Learning for Model Predictive Control

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:40:51.947552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:37:27.658078Z digest=sha256:8f76683c24541920e071771fc701f7fed0fc745ab384cdae5cfe730d37621895

Observation 0fb90d78-f436-4320-a716-5141f08363e2 · inbound

JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning cites this paper.

JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning Temporal Difference Learning for Model Predictive Control

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:37:51.941434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:37:19.404335Z digest=sha256:c69649eafa6bd6dc04b0495faeb372f1eed06735f42ad5993d34b3f2a3ed9604

Observation 5deb802f-3537-44be-a08e-e9e8d8a2486d · inbound

Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry cites this paper.

Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:03:28.775037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T02:03:11.540545Z digest=sha256:99e182d7c2fd54947ecfe1465a7f9bbef8ec4693e9a1561bea4cf4f712068593

Observation 1aad3508-dc94-454c-9196-d633a5efbdb8 · inbound

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making cites this paper.

Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making Temporal Difference Learning for Model Predictive Control

Reference 300

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:59:02.097841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T20:54:31.025488Z digest=sha256:839f244b8b510e7d34944d19bf9c9c856e8c40fd1d158bfdf367fb21c9a33022

Observation 5221a487-d671-40b3-8abe-9c1d1dc5c585 · inbound

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation cites this paper.

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation Temporal Difference Learning for Model Predictive Control

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:01:20.259352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T08:57:19.834179Z digest=sha256:542d1a7df036dbc7134062d297cc89a632ac00b39783afc9747bfcc01a0868d2

Observation 6467fad2-d804-4929-a4f4-601c7fa178d1 · inbound

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization cites this paper.

Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization Temporal Difference Learning for Model Predictive Control

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.393138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:46:27.179341Z digest=sha256:b621a948343d7fadfa0f388364a2af6c4ee74d9c318bb46ab12543ba7881fb01

Observation 3d7662a0-b317-4104-9406-3d811556df83 · inbound

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient cites this paper.

Efficient On-policy Visual-RL via Stochastic Decoupled Policy Gradient Temporal Difference Learning for Model Predictive Control

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:48.589473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T17:34:41.053725Z digest=sha256:4702db59627f2a50f1310b2d35f01e72cc5e873c70df9e95c87fbd35f6a1ac89

Observation 9a77346a-24c1-43bc-920f-d581620dccf6 · inbound

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models cites this paper.

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models Temporal Difference Learning for Model Predictive Control

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.836554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:46:28.469913Z digest=sha256:6b9eebf98e0474df734e361aa70aa194ec32d45e8fde4766013289d2b11ca46b

Observation d0ccad1b-1419-47ed-ab0a-ebab28666818 · inbound

Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization cites this paper.

Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.205977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T22:29:34.616531Z digest=sha256:c81547733686eb68e4e3beaee12ee2bfdbbb5ef015280712ff9c079a0a3e876a

Observation a9e4387f-ba8b-4e27-9052-eb84d5f43e5c · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning Temporal Difference Learning for Model Predictive Control

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:03.124660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:dec6292e49fd2034c4516c07c9827afd01b6ee45b260a00948ca53f623142ebe

Observation 62e5c0bf-1b8e-44de-ad89-bc8f631f0feb · inbound

Solving Markov Decision Processes with Future Information via MPC cites this paper.

Solving Markov Decision Processes with Future Information via MPC Temporal Difference Learning for Model Predictive Control

Reference 84

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:00:06.257841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-25T22:00:02.577707Z digest=sha256:13e0c4ca4e3f606f69249b3ca3425c7a34357c9d5fc07d7c7eea5f09f254b645

Observation 9dddc824-db97-47c5-8e68-30ebc3cdc0da · inbound

Valdi: Value Diffusion World Models cites this paper.

Valdi: Value Diffusion World Models Temporal Difference Learning for Model Predictive Control

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:47:05.609873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-02T15:46:50.507626Z digest=sha256:e57d94d5229516e273c8ce9fa4d99b313d07f085776cf2b668a169a7433a6fda

Observation debe30af-548b-4ee4-8683-a8205d23c246 · inbound

DriftWorld: Fast World Modeling through Drifting cites this paper.

DriftWorld: Fast World Modeling through Drifting Temporal Difference Learning for Model Predictive Control

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:24:04.583755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:24:04.583755Z digest=sha256:f9433cd56d682d8dd20288a231d379d808efc9f663af3025ca5d017a33171645

Observation 4c724199-6b02-4ee6-a461-50feb15d93dc · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models Temporal Difference Learning for Model Predictive Control

Reference 254

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:21.168260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:21.168260Z digest=sha256:03faea034f85b9b90a122e243dcb1a4524839881141a64def329a2215566955c

Observation 42a18446-7aa8-464e-8d94-6ef3f37aafbe · inbound

Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations cites this paper.

Toward Goal-Agnostic Joint-Embedding Predictive Control of Partial Differential Equations Temporal Difference Learning for Model Predictive Control

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T12:24:20.765658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:24:20.765658Z digest=sha256:32c73574ac0ff8c5a1154e69447ce9ae912fd561b6062641504413b1278e0211

Observation 7de8c998-94cb-4d96-98a8-5f39e1582169 · inbound

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies cites this paper.

When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual Auditing of Frozen Locomotion Policies Temporal Difference Learning for Model Predictive Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T06:32:51.940717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:32:51.940717Z digest=sha256:0eaecd611ededc2f00fee616f542bf45064964841cee51d9f3b1f27397be4599

Observation 8fd80f4c-4625-4ccb-bb8f-d3e3a4e7c2e8 · inbound

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models cites this paper.

World Action Planner: Generalizable Decision-Making with Action-Conditioned World Models Temporal Difference Learning for Model Predictive Control

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T04:59:31.588147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T04:59:31.588147Z digest=sha256:9bc3d538915efe991e7ec5b28715599b55ec337c8d607a1cb1bd851705db424a

Observation 3e8be49d-0c0e-4320-9bcf-87242aebceb4 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Temporal Difference Learning for Model Predictive Control

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:32.767376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:32.767376Z digest=sha256:8b6934c7fea04c2724c3ab8bf6b7cda20ecd38575cbbe0eb6c8cedabdfc3caa4