Unregularized centered dueling Q-learning converges under a joint spectral radius condition, with value and advantage acting as distinct gains on common and differential Q components.
hub
Switching in Systems and Control
4 Pith papers cite this work, alongside 6,940 external citations. Polarity classification is still indexing.
hub tools
years
2026 4representative citing papers
A model-free RL method arbitrates between a functional baseline policy and a learning policy, transferring agency over time to yield a standalone policy with high goal-reaching rates and competitive returns on continuous-control tasks.
Zero-shot sim-to-real transfer of independently trained RL policies for cart-pole swing-up and stabilization is achieved via sensitivity-guided domain randomization, linear curriculum learning, and first-order action smoothing with Simulink switching logic.
Backstepping control tracks effective surface area of non-convex satellites for drag-based orbital control, with asymptotic stability proofs and an extension for solar panel exposure.
citing papers explorer
-
Spectral Analysis of Dueling Q-Learning
Unregularized centered dueling Q-learning converges under a joint spectral radius condition, with value and advantage acting as distinct gains on common and differential Q components.
-
An Agency-Transferring Model-Free Policy Enhancement Technique
A model-free RL method arbitrates between a functional baseline policy and a learning policy, transferring agency over time to yield a standalone policy with high goal-reaching rates and competitive returns on continuous-control tasks.
-
Zero-shot Transfer of Reinforcement Learning Control Policies for the Swing-Up and Stabilization of a Cart-Pole System
Zero-shot sim-to-real transfer of independently trained RL policies for cart-pole swing-up and stabilization is achieved via sensitivity-guided domain randomization, linear curriculum learning, and first-order action smoothing with Simulink switching logic.
-
Tracking the Effective Surface Area of Non-Convex Satellites
Backstepping control tracks effective surface area of non-convex satellites for drag-based orbital control, with asymptotic stability proofs and an extension for solar panel exposure.