By dynamically freezing low-rank singular components of diffusion policy weights during training, DRIFT-DAgger cuts training time by roughly 11 to 18 percent while keeping task success near full-rank baselines.
Developing Path Planning with Behavioral Cloning and Proximal Policy Optimization for Path-Tracking and Static Obstacle Nudging
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
In autonomous driving, end-to-end methods utilizing Imitation Learning (IL) and Reinforcement Learning (RL) are becoming more and more common. However, they do not involve explicit reasoning like classic robotics workflow and planning with horizons, resulting in strategies implicit and myopic. In this paper, we introduce a path planning method that uses Behavioral Cloning (BC) for path-tracking and Proximal Policy Optimization (PPO) for static obstacle nudging. It outputs lateral offset values to adjust the given reference waypoints and performs modified path for different controllers. Experimental results show that the algorithm can do path following that mimics the expert performance of path-tracking controllers, and avoid collision to fixed obstacles. The method makes a good attempt at planning with learning-based methods in path planning problems of autonomous driving.
citation-role summary
citation-polarity summary
fields
cs.RO 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Dynamic Rank Adjustment in Diffusion Policies for Efficient and Flexible Training
By dynamically freezing low-rank singular components of diffusion policy weights during training, DRIFT-DAgger cuts training time by roughly 11 to 18 percent while keeping task success near full-rank baselines.