Pith. sign in

Paper Citation Record · LEDGER

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

As of 22 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 27 inbound Pith citation observations for arXiv:2505.22642.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22642 v3

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:08:36.024521Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:22:49.641234Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T17:17:25.644428Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f5493f3c-5de9-48ed-9fd1-5cb4aace6187 · outbound

This paper cites Layer Normalization.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Layer Normalization

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:33.100859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:33.100859Z digest=sha256:bd2b384409df01842232ccd6a220a3011e37ae1e7aad7a57bb5aafcfc455ff72

Observation 1235b156-3a87-4be4-897f-70b61df95cce · outbound

This paper cites Emergence of Locomotion Behaviours in Rich Environments.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Emergence of Locomotion Behaviours in Rich Environments

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.202704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.202704Z digest=sha256:f44dff3f07e4755912515d059888bd235cd142a30267b595eccc64d213f12233

Observation 7c7a0b44-19dc-4d93-b28a-a96577f12dfe · outbound

This paper cites Eureka: Human-Level Reward Design via Coding Large Language Models.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Eureka: Human-Level Reward Design via Coding Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.267219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.267219Z digest=sha256:657c9dc627bf223a74dcc2b25e315bd2dc691c531d4fd2f521687f17c6b23a7b

Observation 0a71d1f0-90ec-4e8e-9066-b8e3a199f01d · outbound

This paper cites Reinforcement learning with action sequence for data-efficient robot learning.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Reinforcement learning with action sequence for data-efficient robot learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.668997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.668997Z digest=sha256:ab28aef977f515bec47ee875b5f55d6a7593784963aed18cd874449306004e75

Observation 2591ff17-1047-4344-9c37-9a96b3937f3d · outbound

This paper cites MuJoCo Playground.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control MuJoCo Playground

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.735651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.735651Z digest=sha256:eca772396a9df10ff5c476d39ab4f10cb9a127e50541ab3097481e64e461e4ff

Observation 2102bee5-da03-4fa1-bbee-ee651aeb7cb4 · outbound

This paper cites TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.794224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.794224Z digest=sha256:a135803125b8da2b666eacfb3259780f33cdab5ae293e778e9bd9d26e2676c43

Observation ee650ec6-90b5-4333-a8f0-1fab41780d05 · outbound

This paper cites We provide learning curves on a 39 tasks from HumanoidBench (Sferrazza et al.,.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control We provide learning curves on a 39 tasks from HumanoidBench (Sferrazza et al.,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:08:36.885361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T13:08:35.903647Z digest=sha256:877ccceb4f41809beb3cbe43dfee25c04dff559809023e7985a8a32b6d802c7d

Observation b9cf310f-1e12-451c-acbb-52133d843f96 · outbound

This paper cites an unresolved cited work.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:08:36.572196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T13:08:35.955548Z digest=sha256:67f8922e632eeeb88c153f161899a0ccb703b5b6967121199fd19e677491a98c

Observation 281762aa-177b-4eb7-86f8-0970e67f2d0e · outbound

This paper cites an unresolved cited work.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:08:36.403039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T13:08:36.024521Z digest=sha256:220aaf5f0d1f88f9d57d72b8c5a4a23352e6b189521a6a1908b958a70d4ab037

Observation aaba6a32-545c-4d45-9701-714c57c5d9bc · outbound

This paper cites Mastering Diverse Domains through World Models.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Mastering Diverse Domains through World Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:34.310435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:34.310435Z digest=sha256:b7f7e6d0a0eee47c9cc504fb9444473a6b9deb22ee018d72f348c8f518f55163

Observation 655e306e-0adf-46f3-bb6b-fb417a5520db · outbound

This paper cites Asymmetric Actor Critic for Image-Based Robot Learning.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Asymmetric Actor Critic for Image-Based Robot Learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.397838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.397838Z digest=sha256:187217bb63e4665ca5d1972430f754ca676655e240869bbe6126115fe3d515e5

Observation 159c74a5-f58e-4a50-b4f0-f18c26aaff1d · outbound

This paper cites Proximal Policy Optimization Algorithms.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Proximal Policy Optimization Algorithms

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.543834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.543834Z digest=sha256:a64c72f19facae359738d29e2701fa4fef7f8980d0ce790022b85b8c3109fcfa

Observation fa909e12-d20a-4823-a413-c388ac9228f6 · outbound

This paper cites Nauman, Michal, Bortkiewicz, Michał, Miło ´s, Piotr, Trzcinski, Tomasz, Ostaszewski, Mateusz, and Cygan, Marek.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Nauman, Michal, Bortkiewicz, Michał, Miło ´s, Piotr, Trzcinski, Tomasz, Ostaszewski, Mateusz, and Cygan, Marek

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:35.336024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:35.336024Z digest=sha256:391658cf504926a2da4fcc6ddfe711b1a0f08a5ecb1c358daffc6eef0fd3d59c

Observation 070edc42-4871-40cc-b768-1ba36e41730a · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Soft Actor-Critic Algorithms and Applications

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:33.335906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:33.335906Z digest=sha256:1e4c42c1f6db57ee7180669ad221782bc9d8e0e9b4b92716d2c12186674adf0e

Observation 3349cc25-8617-4799-b1a8-445fb5b43491 · outbound

This paper cites Simplifying Deep Temporal Difference Learning.

FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Simplifying Deep Temporal Difference Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T13:08:33.176135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:08:33.176135Z digest=sha256:f1b5235a3dba43e84a23bdab82a46dd552550b0cd2214ef7aeab0b5d8f5dfa78

Pith citing papers

Observation 97925b02-2367-49af-a484-f48bd0ff2237 · inbound

ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation cites this paper.

ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:47:13.880014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T09:46:38.742771Z digest=sha256:5a7f3c12fa0589cbae32a5bf5a842da3d46e5332bfd5f13bb78ad0046ec97526

Observation b49bccbb-854c-45b7-846e-dc77347ca93a · inbound

Relative Entropy Pathwise Policy Optimization cites this paper.

Relative Entropy Pathwise Policy Optimization FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:22:04.004454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T04:19:15.018380Z digest=sha256:144295a6a32a86f53165095cd53bf7a5dd6b415c1cf588d44d4f3636831a31cb

Observation fbdc4bdc-a63f-468e-84ca-2fba47c89ba3 · inbound

Flow Matching Policy Gradients cites this paper.

Flow Matching Policy Gradients FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T13:07:09.649092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:07:09.649092Z digest=sha256:87021a29d60c95b7fac61e83d925fd579486c5f2f704941d2a3d1d162687381f

Observation 2866452a-e27c-4565-b71c-001936c8ccbe · inbound

Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies cites this paper.

Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T04:39:07.111116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:39:07.111116Z digest=sha256:c5fda35e185f5c65a8d7ce0840d926f46b6c11caa2a870fe3f753471d1244dbf

Observation 7351708d-747d-4a67-89e9-f54c119cad8a · inbound

Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents cites this paper.

Simplicial Embeddings Improve Sample Efficiency in Actor-Critic Agents FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-04T09:48:07.140767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:48:07.140767Z digest=sha256:94214d938bd95fd8efe2b3a4a56664b117695a119acf5a96c0e6b61363339474

Observation a27929e4-c029-47d7-8d50-eaff230a2b04 · inbound

Stable Deep Reinforcement Learning via Isotropic Gaussian Representations cites this paper.

Stable Deep Reinforcement Learning via Isotropic Gaussian Representations FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T21:42:30.811298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:42:30.811298Z digest=sha256:7e95cbe64c5c553183d7f2b49db7c8c71a753c5274e8f1b9d5088b30692dbe09

Observation 9db694ae-f907-40ab-903c-7895f5901824 · inbound

What Matters for Simulation to Online Reinforcement Learning on Real Robots cites this paper.

What Matters for Simulation to Online Reinforcement Learning on Real Robots FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 61

Resolution
malformed identifier
no resolver link, observed 2026-08-02T21:36:20.383255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:36:20.383255Z digest=sha256:40ce8a2147b82c82db6e3ddffd9c72fba860b8e261b45c3eacb19281532e1a68

Observation a8d28a6b-6fde-4bc7-aad6-60991c46c938 · inbound

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control cites this paper.

FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:20:00.716044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T12:17:59.392158Z digest=sha256:7933bb547c42a89772d8a1e35512395520057110a99bebeb9649acb8edcbda8a

Observation 7965dc9e-b615-46c1-ba23-0db11a6cc1f5 · inbound

ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors cites this paper.

ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:39:55.201294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T09:37:02.168898Z digest=sha256:3f7354d9de0b15d3837f6960b81c93eb0bfb4eb4c925ac081f0a4143bfc83c57

Observation 6fb8192c-e341-4819-8e94-4c1ba6950290 · inbound

Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning cites this paper.

Maximum Entropy Behavior Exploration for Sim2Real Zero-Shot Reinforcement Learning FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T02:32:25.303069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:32:25.303069Z digest=sha256:8e09b6ac2c81bc609e49b4f45e63210e8f3990874d05389aa3f885719a844153

Observation 4ef4803b-6b08-4199-bd95-e1d9845061b8 · inbound

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control cites this paper.

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:15:49.991344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T20:04:56.512544Z digest=sha256:3a422fa5327d72d90be37a3546f26082163c1f93e4034a665c5c0313723f5bbb

Observation 8fc3f8c7-ad06-409c-8f5f-c8b83a3f4394 · inbound

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control cites this paper.

FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 75

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:12:41.347846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T17:08:31.770889Z digest=sha256:d84f3f823827e671c3f49e775275003ca2a73783597eb8cd7b805760117c6705

Observation 859f9511-8d8f-4cc7-bdb3-1dd385583ef7 · inbound

Hyperfastrl: Hypernetwork-based reinforcement learning for unified control of parametric chaotic PDEs cites this paper.

Hyperfastrl: Hypernetwork-based reinforcement learning for unified control of parametric chaotic PDEs FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:25:59.098449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:09:00.155657Z digest=sha256:0b3f1154be03346a88f7cea9af6cef8d0313d7313ef145136ffc841bf332d14f

Observation c79f349e-304d-4150-a272-ec80ac21abbb · inbound

Dyna-Style Safety Augmented Reinforcement Learning: Staying Safe in the Face of Uncertainty cites this paper.

Dyna-Style Safety Augmented Reinforcement Learning: Staying Safe in the Face of Uncertainty FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:36:39.975855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T16:34:33.155628Z digest=sha256:d5d18ec3fe0c8ace72bc2c7758d893d343268fb3a14a1810eb48cdd0ea126fd6

Observation 3f3931e2-3c9e-48d3-80ed-d991a8c0028e · inbound

When Does Non-Uniform Replay Matter in Reinforcement Learning? cites this paper.

When Does Non-Uniform Replay Matter in Reinforcement Learning? FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:36:24.652733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T05:33:19.038889Z digest=sha256:0063096c9be38ff1c8f87d8150b81d27787ff08a52ce0f8c609c5e2b4bef2f10

Observation fff6a116-00a0-407d-9510-49fc3dde82cd · inbound

When Does Non-Uniform Replay Matter in Reinforcement Learning? cites this paper.

When Does Non-Uniform Replay Matter in Reinforcement Learning? FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:32:24.780742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T06:27:38.643667Z digest=sha256:88a858e17299e71630dfb2d73e01ebe39085a4988240765e3ad905eab67ab785

Observation d2f2f081-93ed-46b5-9226-991718f848a6 · inbound

When Does Non-Uniform Replay Matter in Reinforcement Learning? cites this paper.

When Does Non-Uniform Replay Matter in Reinforcement Learning? FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:09:12.684839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T23:04:12.943222Z digest=sha256:bb7d783b74197d954e711b086c4ef252a8a01ac7e49a5756343459a8148ade76

Observation 6a7f2e3b-1d86-4337-a3b3-6cf0714afbee · inbound

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms cites this paper.

UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:13:16.550765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T07:09:19.137932Z digest=sha256:44f43057041bd9e901efb24c7e97cee9477f962cdcfbc39f67b731ce09ac1265

Observation 01d556f2-5392-45c8-acd9-ace790219c8d · inbound

Representation Learning Enables Scalable Multitask Deep Reinforcement Learning cites this paper.

Representation Learning Enables Scalable Multitask Deep Reinforcement Learning FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:46:55.700770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T03:02:59.297911Z digest=sha256:befb160a7cf76893f5368a783a805ec53321812dad07171a4fa50f520d6a7b46

Observation 06805b5e-9603-4903-a27a-f86b69aa352e · inbound

ReFPO: Reflow Regularization for Flow Matching Policy Gradients cites this paper.

ReFPO: Reflow Regularization for Flow Matching Policy Gradients FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:19:37.839390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T14:38:54.754049Z digest=sha256:7714a83e3e9a1ff3c5780bb086a5dcbb9c6a6c3e5ca53aaa357f2049e198e778

Observation c35be60b-c9be-474a-a549-258d5ca42a73 · inbound

AnyBody: Free-Form Whole-Body Humanoid Control from Arbitrary Keypoint Guidance cites this paper.

AnyBody: Free-Form Whole-Body Humanoid Control from Arbitrary Keypoint Guidance FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:04:28.870067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T07:44:51.002937Z digest=sha256:3f4e9458eb90b1a0198999c093e0864c6bdb2c7fcbbd01a5a8f161ea4d11803a

Observation 6fa03dd2-388b-4323-b04e-84fffc7d302a · inbound

ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning cites this paper.

ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 100

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T08:24:26.913170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-30T06:06:42.672173Z digest=sha256:fdf2116638ab99bb6b38a96c3ff885e3a83b3ecb029b44476a0b3960bdd66e03

Observation 777d2734-968a-4372-809d-a0cc91583f46 · inbound

FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion cites this paper.

FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:42.276740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-01T05:28:28.298787Z digest=sha256:e4820434da902e7eec5f09462b112565b6aacea0ac3c7d5f90271bad1860b571

Observation 1864b3a0-c77c-4262-af4b-0169085e9922 · inbound

Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains cites this paper.

Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-07-10T17:17:25.647597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T17:11:13.461970Z digest=sha256:07cc27ee0d349339162d13fbc34444d6e290ca936d7e62b598fc8720e55fbc05

Observation 60a3ca50-70f6-42c6-bbec-3d53dd60d464 · inbound

Scaling Behavior Foundation Model for Humanoid Robots cites this paper.

Scaling Behavior Foundation Model for Humanoid Robots FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-02T00:03:18.235978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:03:18.235978Z digest=sha256:4c3fe98adc45235d14ca493bb002d0199d5ed97e8a691ec6ae896106a5ae9c8f

Observation d1335718-9dcf-4d1a-90b1-8267695f4100 · inbound

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts cites this paper.

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T00:17:16.510189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:17:16.510189Z digest=sha256:d43d6b30ada5cee774909e5f94830bf7b2d23570e6d2a5361494179fe88546db

Observation 55562711-9a6d-445e-a4e3-fe6f8c32405a · inbound

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL cites this paper.

Learning Loco-Manipulation From SMPC Demonstrations With Sparse Offline-to-Online RL FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T00:22:49.641234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:22:49.641234Z digest=sha256:0c4fca7bc278d3e0f30b5012694e4cf87157882d02863daa8e4ee4fedcb71633