Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:27:30.592023Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 0 inbound Pith citation observations for arXiv:2506.12095.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:27:30.592023Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e260941e-9d6f-4510-b2ae-7ccda0e66247 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Learning humanoid locomotion with transformers,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7889185-0ee8-4d79-918b-ee4f71d97b26 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Real-world humanoid locomotion with reinforcement learning,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 52a8d7ac-5b49-41e5-8d74-14e8d27525b2 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Model predictive control,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f4bab07a-8f8e-4064-9cf6-977f1283d99c · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Aleatoric and epistemic uncertainty with random forests,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87224d6d-60e6-4b83-a8dd-19afe8e017c6 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Aleatoric and epistemic uncertainty in machine learning: An introduction to concepts and methods,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d584cc8-db7c-4a6e-81f4-d036127ff71f · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Stochasticity in Motion: An Information-Theoretic Approach to Trajectory Prediction
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b5c315c-d480-435d-84ac-a50798c6409d · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Learning Through Retrospection: Improving Trajectory Prediction for Automated Driving with Error Feedback
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 878a5feb-84b3-4579-beea-fa4067b7500c · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Temporal difference learning for model predictive control,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e2d9b79-dbe9-4b1f-9065-b6a28d55aaa9 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Td-mpc2: Scalable, robust world models for continuous control,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56bbb6f7-6191-4408-94d9-414ebade58b5 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Deep reinforce- ment learning in a handful of trials using probabilistic dynamics models,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70dd384c-1d61-43f9-ac3f-887c72d72b08 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Model-Based Offline Planning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a772cadb-ca4b-44c8-844c-52e155ff9de9 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion TD-M(PC)$^2$: Improving Temporal Difference MPC Through Policy Constraint
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986fb49b-5f58-42ba-9610-d3c7d1488490 · outbound
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cdcabf3-1b02-4939-94c8-9de1834b17a6 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b14b05d-c8fa-4c20-b22a-d15cb39aa7d0 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c337b9cf-f2b7-4f20-9d21-9d6a3f968b11 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Reinforcement learning for humanoid robotics,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 11b382bb-3341-4f1c-90fb-c1db1f7a0b97 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Learning off-policy with online planning,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f123cf6a-1eea-4d54-8c5a-9b4320964f46 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conformal prediction in manifold learning,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 614f192f-7e1f-4661-83c0-acde746bf96c · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conformal Prediction with Learned Features
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d667aa7-e41b-4da9-884c-c297c762fdd9 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Confor- mal prediction for semantically-aware autonomous perception in urban environments,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3a52c1e-a316-4937-83dc-e277ce094e56 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conformal prediction for uncertainty-aware planning with diffusion dynamics model,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 07eba7e6-feff-4612-8708-cba7edc3e584 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Adaptive conformal prediction for motion planning among dynamic agents,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82632f0d-106c-4c04-aac0-bbe697569ced · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Safe planning in dynamic environments using conformal prediction,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6f144c20-bd22-4315-af55-7676327a4b1d · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Safe perception-based control under stochastic sensor uncertainty using con- formal prediction,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 94498e85-f25f-4691-9792-b54b2eee61c2 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conformal decision theory: Safe autonomous decisions from imperfect predictions,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 668e69c2-056f-45a2-9bbf-c26ed3c50442 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Safe pomdp online planning among dynamic agents via adaptive conformal prediction,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9982d18c-a52c-4002-948c-34f7c9562c1e · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conformal policy learning for sensorimotor control under distribution shifts,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 45659071-8e67-416d-9c9d-f0cb3d6260ba · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conformalized Teleoperation: Confidently Mapping Human Inputs to High-Dimensional Robot Actions
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 129a9df3-0ca0-4cd3-9f37-ab4523e42d85 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Stabilizing off- policy q-learning via bootstrapping error reduction,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8364e28-c826-42d7-b104-9488c5a1d8bd · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Conservative q-learning for offline reinforcement learning,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03d3fb43-b9bb-442d-b1ed-df58065d65db · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion A minimalist approach to offline reinforce- ment learning,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0913dee-c670-42da-9b0f-7160d230e533 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Off-policy deep reinforcement learning without exploration,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f69c47d4-639e-4bb8-a9b2-5ee79a991814 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b556c63-0a75-4c9d-a391-14da58e48501 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Extreme Q-Learning: MaxEnt RL without Entropy
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05760fd6-b9e7-4ee7-9971-fb4e4a86156d · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Offline Reinforcement Learning with Implicit Q-Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9f23434-f2ee-4508-8fb0-c5c7e8889472 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Aggressive driving with model predictive path integral control,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8df83b-7ac8-4750-ad88-c1415d4c9e99 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Trust region policy optimization,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc279f4d-39da-4c85-b56e-c235f4528584 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f335dd01-54b4-4e0e-88e6-d9ab13ca0f86 · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion Imitation is not enough: Ro- bustifying imitation with reinforcement learning for challenging driving scenarios,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a59ac21-e20c-4335-982e-fa8a68de9b1a · outbound
DoublyAware: Dual Planning and Policy Awareness for Temporal Difference Learning in Humanoid Locomotion AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.