RBWE applies offline RL with Q-ensembles and Gaussian mixture policies to bandwidth estimation, cutting overestimation errors by 18% and raising the 10th percentile QoE by 18.6% over GCC.
From Ember to Blaze: Swift Interactive Video Adaptation via Meta- Reinforcement Learning,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.SY 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Robust Bandwidth Estimation for Real-Time Communication with Offline Reinforcement Learning
RBWE applies offline RL with Q-ensembles and Gaussian mixture policies to bandwidth estimation, cutting overestimation errors by 18% and raising the 10th percentile QoE by 18.6% over GCC.