BIOMAP achieves the optimal reward on the Mask Cliff Walking benchmark by reconstructing the state graph from action vectors, but this hinges on an unstated assumption that states are geometric positions.
Belief space planning assuming maximum likelihood observations
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
eess.SY 1years
2024 1verdicts
REJECT 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
A Model-free Biomimetics Algorithm for Deterministic Partially Observable Markov Decision Process
BIOMAP achieves the optimal reward on the Mask Cliff Walking benchmark by reconstructing the state graph from action vectors, but this hinges on an unstated assumption that states are geometric positions.