Policies trained under stationary latent ambiguity, implemented by refreshing the latent parameter, preserve robustness to regime shifts better than policies trained under a fixed latent draw.
Proposition(Doob’s theorem).Let X and Y be Polish spaces, equipped with their Borel σ-algebras
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Robust Control under Stationary Ambiguity
Policies trained under stationary latent ambiguity, implemented by refreshing the latent parameter, preserve robustness to regime shifts better than policies trained under a fixed latent draw.