Pith. sign in

Online algorithms for POMDPs with continuous state, action, and observation spaces

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Online solvers for partially observable Markov decision processes have been applied to problems with large discrete state spaces, but continuous state, action, and observation spaces remain a challenge. This paper begins by investigating double progressive widening (DPW) as a solution to this challenge. However, we prove that this modification alone is not sufficient because the belief representations in the search tree collapse to a single particle causing the algorithm to converge to a policy that is suboptimal regardless of the computation time. This paper proposes and evaluates two new algorithms, POMCPOW and PFT-DPW, that overcome this deficiency by using weighted particle filtering. Simulation results show that these modifications allow the algorithms to be successful where previous approaches fail.

fields

cs.AI 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

citing papers explorer

Showing 1 of 1 citing paper.

  • LLM-Guided Probabilistic Program Induction for POMDP Model Estimation cs.AI · 2025-05-04 · conditional · none · ref 32 · internal anchor

    LLM-guided probabilistic program induction can learn low-complexity POMDP models from ten demonstrations and outperform tabular learning, behavior cloning, and direct LLM planning in simulated and real robot domains.