D3HRL combines distributed causal discovery over multiple time lags with conditional independence testing to learn delayed action effects and build hierarchical policies in long-horizon tasks.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
D3HRL: A Distributed Hierarchical Reinforcement Learning Approach Based on Causal Discovery and Spurious Correlation Detection
D3HRL combines distributed causal discovery over multiple time lags with conditional independence testing to learn delayed action effects and build hierarchical policies in long-horizon tasks.