Pith. sign in

Paper Citation Record · LEDGER

NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2305.16986.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.16986 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:51:50.963143Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:09:37.364862Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7082600f-ae2b-4568-b8da-8fcb1d8c980f · inbound

Agent AI: Surveying the Horizons of Multimodal Interaction cites this paper.

Agent AI: Surveying the Horizons of Multimodal Interaction NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 112

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:25:59.318626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T14:25:58.876978Z digest=sha256:76e4d268d21b297d6f2cc4a4067dbd9c7a32c3bcf28cdfc84870fcf17c184635

Observation ded1a298-dcda-49b3-9238-3e19cb415de0 · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:20.565357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:4f0ef20342d3680d1145d6482b066726300a4a42f99970e31dc05a4338b80388

Observation f2caa80c-0171-4fd1-a0f7-bb98bd334d8c · inbound

Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks cites this paper.

Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:51:36.388976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T19:51:36.137985Z digest=sha256:00e1e32caa97858036335af6ff73c7c76ab48131e45c0b5acf860230589211ec

Observation 11c3bad6-0b8d-4f07-b6de-e5207a512cf4 · inbound

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization cites this paper.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:50.963143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:50.963143Z digest=sha256:f5def470a30176f17a60ac83fb6888d53212e74edd0192afd5dc5e68c1014bf2

Observation c91a7d8e-19bf-4711-9c03-293b3c2b4026 · inbound

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling cites this paper.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:05.744748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:05.744748Z digest=sha256:a96390cdbd65e87c8ca0ea48097b4c3b7e1eb6f0cc27bd951bc2282726431dc1

Observation 8c135422-e0dd-4747-b58c-1a57e16c7284 · inbound

MSNav: Zero-Shot Vision-and-Language Navigation with Dynamic Memory and LLM Spatial Reasoning cites this paper.

MSNav: Zero-Shot Vision-and-Language Navigation with Dynamic Memory and LLM Spatial Reasoning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T18:39:12.470115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:39:12.470115Z digest=sha256:37abadfc105e84eeeed2207f13e55a14995c0dc7468491bbac0c5642965679de

Observation f3422f1a-3965-49c3-afdf-8a41250b5b9e · inbound

DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation cites this paper.

DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:00:45.037296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:00:45.037296Z digest=sha256:6f26f425d65ebca81881344604175576e29d0a0a41a341fd80a11bc4cfbcc076

Observation 0017e3d3-9ebd-477e-9b16-c90f2ae7cbed · inbound

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning cites this paper.

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T12:37:56.580627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:37:56.580627Z digest=sha256:95729549330d7936c39f8b76369c9541eac84eb4407b052fa2455ee8a23b319b

Observation 6799a38b-e08b-432e-bf85-ed461025272c · inbound

Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation cites this paper.

Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:11:21.132290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T23:10:37.814878Z digest=sha256:f0716d5f1d0e8758a6476a0dd63e8843acc0c14fdf7af8c865ca4c3de95dec6b

Observation 1a61d812-a500-4d10-ade0-c502d2c6efc1 · inbound

SG-CoT: An Ambiguity-Aware Robotic Planning Framework using Scene Graph Representations cites this paper.

SG-CoT: An Ambiguity-Aware Robotic Planning Framework using Scene Graph Representations NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T17:57:56.499039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:57:56.499039Z digest=sha256:a32fdee41dbb2a1726e4409311bd44d7993e481e592c234cdb759d383ee8ea95

Observation 97343341-7e49-46fa-8f53-2dc94ba54834 · inbound

How Far Are Large Multimodal Models from Human-Level Spatial Action? A Benchmark for Goal-Oriented Embodied Navigation in Urban Airspace cites this paper.

How Far Are Large Multimodal Models from Human-Level Spatial Action? A Benchmark for Goal-Oriented Embodied Navigation in Urban Airspace NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:25:56.913695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:09:39.510770Z digest=sha256:6e06d01f0d7050839602b38b064c25d4edc1467d8f20e0e429b615c560528524

Observation a60f1f0c-baf3-4cf9-b30f-e26ee35b056c · inbound

PhotoFlow: Agentic 3D Virtual Photography Missions cites this paper.

PhotoFlow: Agentic 3D Virtual Photography Missions NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.046941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T04:33:24.355622Z digest=sha256:5a55d1a63abd3b6f22ac46371f43fe9de009ff9ad9f1ba9769c3ed8801448ea5

Observation 0ec91a74-4cc8-4623-834b-22bcaeccbb31 · inbound

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning cites this paper.

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.323474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T22:23:24.536258Z digest=sha256:91e5a718c4a20a5beac8de0800a84451398d55d40890654a7ea58ca61125bde5

Observation 4b936c32-405f-478c-af5c-add0ac2e8750 · inbound

BIT-Nav: Brain-Inspired Trajectory Memory for Embodied Navigation cites this paper.

BIT-Nav: Brain-Inspired Trajectory Memory for Embodied Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:09:37.366333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T13:52:31.453516Z digest=sha256:bf7288450a1d87f110faa7626f0d30bba29f35de13b29c80d47aee66fc8d806e

Observation 87951153-cd2e-4ee6-a848-7ac3792bcd51 · inbound

ReferTrack: Referring Then Tracking for Embodied Visual Tracking cites this paper.

ReferTrack: Referring Then Tracking for Embodied Visual Tracking NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T11:00:24.420418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:00:24.420418Z digest=sha256:3e1612614e83e129c90ba3a53dbb5be1e8f094ed509458f3dd6d9fd3cfd06354