RE-SAC disentangles aleatoric and epistemic risks via IPM regularization on the critic and a diversified Q-ensemble, yielding higher rewards and lower estimation error than vanilla SAC in simulated bus corridor control.
A headway-based approach to eliminate bus bunching: Systematic analysis and comparisons,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2026 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
RE-SAC: Disentangling aleatoric and epistemic risks in bus fleet control: A stable and robust ensemble DRL approach
RE-SAC disentangles aleatoric and epistemic risks via IPM regularization on the critic and a diversified Q-ensemble, yielding higher rewards and lower estimation error than vanilla SAC in simulated bus corridor control.