Pith. sign in

Robotic Test Tube Rearrangement Using Combined Reinforcement Learning and Motion Planning

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

A combined task-level reinforcement learning and motion planning framework is proposed in this paper to address a multi-class in-rack test tube rearrangement problem. At the task level, the framework uses reinforcement learning to infer a sequence of swap actions while ignoring robotic motion details. At the motion level, the framework accepts the swapping action sequences inferred by task-level agents and plans the detailed robotic pick-and-place motion. The task and motion-level planning form a closed loop with the help of a condition set maintained for each rack slot, which allows the framework to perform replanning and effectively find solutions in the presence of low-level failures. Particularly for reinforcement learning, the framework leverages a distributed deep Q-learning structure with the Dueling Double Deep Q Network (D3QN) to acquire near-optimal policies and uses an A${}^\star$-based post-processing technique to amplify the collected training data. The D3QN and distributed learning help increase training efficiency. The post-processing helps complete unfinished action sequences and remove redundancy, thus making the training data more effective. We carry out both simulations and real-world studies to understand the performance of the proposed framework. The results verify the performance of the RL and post-processing and show that the closed-loop combination improves robustness. The framework is ready to incorporate various sensory feedback. The real-world studies also demonstrated the incorporation.

citation-role summary

background 1

citation-polarity summary

fields

cs.RO 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Mobile Manipulation Planning for Tabletop Rearrangement

cs.RO · 2025-05-24 · conditional · novelty 6.0

STRAP V2 lets a mobile manipulator perform multiple pick-and-place operations from one base position and uses state re-exploration to reduce total cost and planning time in simulated tabletop rearrangement.

citing papers explorer

Showing 1 of 1 citing paper.

  • Mobile Manipulation Planning for Tabletop Rearrangement cs.RO · 2025-05-24 · conditional · none · ref 15 · internal anchor

    STRAP V2 lets a mobile manipulator perform multiple pick-and-place operations from one base position and uses state re-exploration to reduce total cost and planning time in simulated tabletop rearrangement.