Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T11:20:20.073159Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2608.09762.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T11:20:20.073159Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
31 of 31 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7953fa82-b810-499b-bb1c-efcf13a82c01 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition A review of learning-based dynamics models for robotic manipulation |Science Robotics
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f175a95-4231-4d2c-8106-761b11c7705a · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Learning Con- tinuous Control Actions for Robotic Grasping with Reinforcement Learning,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d0cd25a3-7530-46ae-af01-50fc90390d03 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition General- Purpose Sim2Real Protocol for Learning Contact-Rich Manipulation With Marker-Based Visuotactile Sensors,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3eac9bce-f94b-4229-a64e-e3e40b92467d · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13405272-286b-4dc9-b277-7b1d07376018 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Srl-vic: A variable stiffness-based safe reinforcement learning for contact-rich robotic tasks,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 95e6ccef-cb1e-475c-8382-cf36fc8ff216 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Deep reinforcement learning for robotics: A survey of real-world successes,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee437a5f-e70d-4f2a-be11-d2916a6e6eee · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Improving vision-language-action model with online reinforcement learning,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ae2eab6f-486a-4215-bef2-fd81a99c225d · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Hi-WM: Human-in-the-World-Model for Scalable Robot Post-Training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bab35a4-89e1-4c86-b096-dcf22b84ad41 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Rl-100: Performant robotic manipulation with real-world reinforcement learning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96e95e0a-14ed-41b8-afe7-ff5c1cc4a317 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Rlinf-user: A unified and extensible system for real-world online policy learning in embodied ai,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dca26e27-0d33-4020-bab7-77c33600fa49 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation With Large Language Models,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d314379b-deee-43a2-aa96-5c385f267da2 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Serl: A software suite for sample- efficient robotic reinforcement learning,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6fd6e15-207f-490b-b2ad-28e8b8d3db58 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Precise and dexterous robotic manipulation via human-in-the-loop reinforcement learning,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fe47e55-fee9-49fc-b126-586f27ee244d · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22e82f08-1792-4285-9c83-703042f22bcf · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Reinflow: Fine-tuning flow matching policy with online reinforcement learning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 61f3d213-e333-428e-8051-c675de64a402 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Hybrid reward architecture for reinforcement learning,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0142220e-b7b5-49fc-8fe4-3080dc7f2d20 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Efficient online reinforcement learning with offline data,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df969530-88a0-45bf-8ae8-eb2b1f682f6d · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Deep Reinforcement Learning in Parameterized Action Space
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8be1923-6991-4df6-864a-890db2122199 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe4d146d-8b36-413d-9a1c-539d15bd0047 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Multi-Pass Q-Networks for Deep Reinforcement Learning with Parameterised Action Spaces
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a078d8e-7ac0-4a00-8fc4-eb2c7f5db353 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Multi-agent actor-critic for mixed cooperative-competitive environments,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36ffc7a6-71f5-4857-84c9-17b733237f35 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition The surprising effectiveness of ppo in cooperative multi-agent games,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55b51ff4-6707-4e05-9489-6b5836aefa71 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Asynchronous actor-critic for multi-agent reinforcement learning,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5d5e5942-bb6d-414a-812b-bd687254c0a5 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Action decoupled sac reinforcement learning with discrete-continuous hybrid action spaces,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 83659ad8-fe57-4c94-b34d-b03c64bccc8a · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Effective multi-agent deep reinforcement learning control with relative entropy regularization,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e5eeae7a-6621-4cf4-a573-156865f8c01e · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Soft Actor-Critic Algorithms and Applications
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2cabc70-02ae-44eb-b8ae-be067ec5d036 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Deep residual learning for image recognition,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5074cb71-2bf8-4341-9c92-132d71c9fe2e · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 641735f2-9b57-48e6-98be-efa0f519e538 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Soft Actor-Critic for Discrete Action Settings
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c43875eb-ce6b-4761-a192-06de815cc0a2 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition A high-force gripper with embedded multimodal sensing for powerful and perception driven grasping,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 80032bd8-1fab-403b-ad4b-09aa3173a7e7 · outbound
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition Root mean square layer normalization,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.