Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:08:36.024521Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 27 inbound Pith citation observations for arXiv:2505.22642.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:08:36.024521Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:22:49.641234Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T17:17:25.644428Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f5493f3c-5de9-48ed-9fd1-5cb4aace6187 · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Layer Normalization
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1235b156-3a87-4be4-897f-70b61df95cce · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Emergence of Locomotion Behaviours in Rich Environments
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c7a0b44-19dc-4d93-b28a-a96577f12dfe · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Eureka: Human-Level Reward Design via Coding Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a71d1f0-90ec-4e8e-9066-b8e3a199f01d · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Reinforcement learning with action sequence for data-efficient robot learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2591ff17-1047-4344-9c37-9a96b3937f3d · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control MuJoCo Playground
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2102bee5-da03-4fa1-bbee-ee651aeb7cb4 · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee650ec6-90b5-4333-a8f0-1fab41780d05 · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control We provide learning curves on a 39 tasks from HumanoidBench (Sferrazza et al.,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b9cf310f-1e12-451c-acbb-52133d843f96 · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 281762aa-177b-4eb7-86f8-0970e67f2d0e · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation aaba6a32-545c-4d45-9701-714c57c5d9bc · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Mastering Diverse Domains through World Models
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 655e306e-0adf-46f3-bb6b-fb417a5520db · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Asymmetric Actor Critic for Image-Based Robot Learning
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 159c74a5-f58e-4a50-b4f0-f18c26aaff1d · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Proximal Policy Optimization Algorithms
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa909e12-d20a-4823-a413-c388ac9228f6 · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Nauman, Michal, Bortkiewicz, Michał, Miło ´s, Piotr, Trzcinski, Tomasz, Ostaszewski, Mateusz, and Cygan, Marek
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 070edc42-4871-40cc-b768-1ba36e41730a · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Soft Actor-Critic Algorithms and Applications
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3349cc25-8617-4799-b1a8-445fb5b43491 · outbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Simplifying Deep Temporal Difference Learning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97925b02-2367-49af-a484-f48bd0ff2237 · inbound
ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b49bccbb-854c-45b7-846e-dc77347ca93a · inbound
Relative Entropy Pathwise Policy Optimization FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation fbdc4bdc-a63f-468e-84ca-2fba47c89ba3 · inbound
Flow Matching Policy Gradients FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2866452a-e27c-4565-b71c-001936c8ccbe · inbound
Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7351708d-747d-4a67-89e9-f54c119cad8a · inbound
Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a27929e4-c029-47d7-8d50-eaff230a2b04 · inbound
Stable Deep Reinforcement Learning via Isotropic Gaussian Representations FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9db694ae-f907-40ab-903c-7895f5901824 · inbound
What Matters for Simulation to Online Reinforcement Learning on Real Robots FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8d28a6b-6fde-4bc7-aad6-60991c46c938 · inbound
FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7965dc9e-b615-46c1-ba23-0db11a6cc1f5 · inbound
ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6fb8192c-e341-4819-8e94-4c1ba6950290 · inbound
Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ef4803b-6b08-4199-bd95-e1d9845061b8 · inbound
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8fc3f8c7-ad06-409c-8f5f-c8b83a3f4394 · inbound
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 859f9511-8d8f-4cc7-bdb3-1dd385583ef7 · inbound
Hyperfastrl: Hypernetwork-based reinforcement learning for unified control of parametric chaotic PDEs FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c79f349e-304d-4150-a272-ec80ac21abbb · inbound
Dyna-Style Safety Augmented Reinforcement Learning: Staying Safe in the Face of Uncertainty FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3f3931e2-3c9e-48d3-80ed-d991a8c0028e · inbound
When Does Non-Uniform Replay Matter in Reinforcement Learning? FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation fff6a116-00a0-407d-9510-49fc3dde82cd · inbound
When Does Non-Uniform Replay Matter in Reinforcement Learning? FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d2f2f081-93ed-46b5-9226-991718f848a6 · inbound
When Does Non-Uniform Replay Matter in Reinforcement Learning? FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6a7f2e3b-1d86-4337-a3b3-6cf0714afbee · inbound
UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 01d556f2-5392-45c8-acd9-ace790219c8d · inbound
Representation Learning Enables Scalable Multitask Deep Reinforcement Learning FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 06805b5e-9603-4903-a27a-f86b69aa352e · inbound
ReFPO: Reflow Regularization for Flow Matching Policy Gradients FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c35be60b-c9be-474a-a549-258d5ca42a73 · inbound
AnyBody: Free-Form Whole-Body Humanoid Control from Arbitrary Keypoint Guidance FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6fa03dd2-388b-4323-b04e-84fffc7d302a · inbound
ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 777d2734-968a-4372-809d-a0cc91583f46 · inbound
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 1864b3a0-c77c-4262-af4b-0169085e9922 · inbound
Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 60a3ca50-70f6-42c6-bbec-3d53dd60d464 · inbound
Scaling Behavior Foundation Model for Humanoid Robots FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1335718-9dcf-4d1a-90b1-8267695f4100 · inbound
LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55562711-9a6d-445e-a4e3-fe6f8c32405a · inbound
Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.