pith. sign in

arxiv: 2109.02066 · v2 · pith:V7EVH52Lnew · submitted 2021-09-05 · 💻 cs.CV

Hierarchical Object-to-Zone Graph for Object Navigation

classification 💻 cs.CV
keywords agentgraphobjectzonenodesreal-timeaccordingaction
0
0 comments X
read the original abstract

The goal of object navigation is to reach the expected objects according to visual information in the unseen environments. Previous works usually implement deep models to train an agent to predict actions in real-time. However, in the unseen environment, when the target object is not in egocentric view, the agent may not be able to make wise decisions due to the lack of guidance. In this paper, we propose a hierarchical object-to-zone (HOZ) graph to guide the agent in a coarse-to-fine manner, and an online-learning mechanism is also proposed to update HOZ according to the real-time observation in new environments. In particular, the HOZ graph is composed of scene nodes, zone nodes and object nodes. With the pre-learned HOZ graph, the real-time observation and the target goal, the agent can constantly plan an optimal path from zone to zone. In the estimated path, the next potential zone is regarded as sub-goal, which is also fed into the deep reinforcement learning model for action prediction. Our methods are evaluated on the AI2-Thor simulator. In addition to widely used evaluation metrics SR and SPL, we also propose a new evaluation metric of SAE that focuses on the effective action rate. Experimental results demonstrate the effectiveness and efficiency of our proposed method.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Light cone QCD sum rules study of the rare radiative $\Xi^{*}_{bb}\to\Xi_b\gamma$ decay

    hep-ph 2026-01 unverdicted novelty 5.0

    Light cone QCD sum rules produce the form factors T1V(q²=0), T2V(q²=0), T1A(q²=0), T2A(q²=0) for Ξ*bb → Ξb γ, from which the decay width is calculated.