Pith. sign in

Paper Citation Record · LEDGER

NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2305.16986.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.16986 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:06:08.513280Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T07:09:37.364862Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7082600f-ae2b-4568-b8da-8fcb1d8c980f · inbound

Agent AI: Surveying the Horizons of Multimodal Interaction cites this paper.

Agent AI: Surveying the Horizons of Multimodal Interaction NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 112

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T14:25:59.318626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-18T14:25:58.876978Z digest=sha256:39efee338b7a9cad98f8d9b7d9bcf60c4568c67dd8d1207f132197812bb2d117

Observation ded1a298-dcda-49b3-9238-3e19cb415de0 · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:20.565357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:606452728669f8cfb6ec0d6ba2314ae429c4c8e42377363a29d63d76991d7a4c

Observation 018b6d5a-777a-4d61-b9be-40de66b54966 · inbound

NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation cites this paper.

NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:19.646539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:19.646539Z digest=sha256:1c59d17efa76f5342e9e0d506fc0a6f56d60d78ea31ccefc449ad35cd2c4e064

Observation f2caa80c-0171-4fd1-a0f7-bb98bd334d8c · inbound

Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks cites this paper.

Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:51:36.388976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T19:51:36.137985Z digest=sha256:b8a481a752a9913f2b9127109a2b26f953888b5f6a7332821a609e7a8610a42d

Observation 8860a15d-ad46-4d1b-9f08-5cf92ced8fb0 · inbound

Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models cites this paper.

Language and Planning in Robotic Navigation: A Multilingual Evaluation of State-of-the-Art Models NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T21:47:59.021845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:47:59.021845Z digest=sha256:459415b6302186e6e1cbaab8b2bae1d9c1ec3add29cf7485701d53f84da9c380

Observation 680dd3e3-6a97-4acd-81c5-c69ce9912374 · inbound

Think Hierarchically, Act Dynamically: Hierarchical Multi-modal Fusion and Reasoning for Vision-and-Language Navigation cites this paper.

Think Hierarchically, Act Dynamically: Hierarchical Multi-modal Fusion and Reasoning for Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T11:06:08.513280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:06:08.513280Z digest=sha256:cf15f9ffccf69cf801cd8c5c64568b7318fccab468b88073038856d041c7ad92

Observation bf7f1fdf-69eb-4300-bdef-3475344fac57 · inbound

LogisticsVLN: Vision-Language Navigation For Low-Altitude Terminal Delivery Based on Agentic UAVs cites this paper.

LogisticsVLN: Vision-Language Navigation For Low-Altitude Terminal Delivery Based on Agentic UAVs NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:55:39.868859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:55:39.868859Z digest=sha256:f63ee9619b6fd9640f76b454df3aaa61e222840837b99cb91b108cef6015dc5e

Observation 8f2a9eee-70ce-4a39-81d1-78360dd66e04 · inbound

VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning cites this paper.

VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T19:13:35.468781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:13:35.468781Z digest=sha256:d8cf252115b6ab32ff9bd938d4b3b834356ad33497dcad7000af90b9a60e8bec

Observation 11c3bad6-0b8d-4f07-b6de-e5207a512cf4 · inbound

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization cites this paper.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:50.963143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:50.963143Z digest=sha256:ef220994c8733f6d3b7ca72fe7adfeb72e35616016707d1798bf8f879bafdb47

Observation c91a7d8e-19bf-4711-9c03-293b3c2b4026 · inbound

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling cites this paper.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:05.744748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:05.744748Z digest=sha256:792337bcdde162e3f188ed1a821130f47a35acf16feb97d90ae772a1aae1aeee

Observation 8c135422-e0dd-4747-b58c-1a57e16c7284 · inbound

MSNav: Zero-Shot Vision-and-Language Navigation with Dynamic Memory and LLM Spatial Reasoning cites this paper.

MSNav: Zero-Shot Vision-and-Language Navigation with Dynamic Memory and LLM Spatial Reasoning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T18:39:12.470115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:39:12.470115Z digest=sha256:f4958e0f9360ddd5486ad2818e944da84c6ff422a2e117c52beb85ca893a535d

Observation 4bc36bf5-c073-4502-831d-1a83b8b0863a · inbound

GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation cites this paper.

GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:57:22.854536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:57:22.854536Z digest=sha256:1ca6eb4ff1e190dfbd8af7d703cc7dbea253c0a925c41b9417e32e6bddcd2252

Observation f3422f1a-3965-49c3-afdf-8a41250b5b9e · inbound

DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation cites this paper.

DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:00:45.037296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:00:45.037296Z digest=sha256:ed09b1aa97114eecc7f6a6489876ebe69f380755ea124ddd9eb50f4c7e8bf157

Observation 0017e3d3-9ebd-477e-9b16-c90f2ae7cbed · inbound

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning cites this paper.

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T12:37:56.580627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:37:56.580627Z digest=sha256:7758360ed307698648bed09da5c12337e45479e39b3ada95b8063c40538410af

Observation 6799a38b-e08b-432e-bf85-ed461025272c · inbound

Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation cites this paper.

Floorplan2Guide: LLM-Guided Floorplan Parsing for BLV Indoor Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:11:21.132290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T23:10:37.814878Z digest=sha256:1d97cee6b2fd1f123ae498bca4081a27c55ca4387a1f23f33f71160a4fe15b26

Observation 1a61d812-a500-4d10-ade0-c502d2c6efc1 · inbound

SG-CoT: An Ambiguity-Aware Robotic Planning Framework using Scene Graph Representations cites this paper.

SG-CoT: An Ambiguity-Aware Robotic Planning Framework using Scene Graph Representations NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T17:57:56.499039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:57:56.499039Z digest=sha256:227fc5cc04fe8ff08346da91cc9c9e977ce076f5cbd9c7794be7dc4b61ba8210

Observation 97343341-7e49-46fa-8f53-2dc94ba54834 · inbound

How Far Are Large Multimodal Models from Human-Level Spatial Action? A Benchmark for Goal-Oriented Embodied Navigation in Urban Airspace cites this paper.

How Far Are Large Multimodal Models from Human-Level Spatial Action? A Benchmark for Goal-Oriented Embodied Navigation in Urban Airspace NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:25:56.913695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T18:09:39.510770Z digest=sha256:51a5f97a9db73f23f0b398a24022a296ba5ffa7a0e90e8375ac78fdc4fc4f50f

Observation a60f1f0c-baf3-4cf9-b30f-e26ee35b056c · inbound

PhotoFlow: Agentic 3D Virtual Photography Missions cites this paper.

PhotoFlow: Agentic 3D Virtual Photography Missions NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:35:21.046941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T04:33:24.355622Z digest=sha256:9b72839b5525a45d57b3c19fa692406c17e3d0581ba11cbfc169ff7542fe5ce7

Observation 0ec91a74-4cc8-4623-834b-22bcaeccbb31 · inbound

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning cites this paper.

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.323474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T22:23:24.536258Z digest=sha256:7e2b099c4cbb4c0dbf45e495147e0f4883c04345c2153ffa86e3d1ec3123b2f3

Observation 4b936c32-405f-478c-af5c-add0ac2e8750 · inbound

BIT-Nav: Brain-Inspired Trajectory Memory for Embodied Navigation cites this paper.

BIT-Nav: Brain-Inspired Trajectory Memory for Embodied Navigation NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:09:37.366333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T13:52:31.453516Z digest=sha256:370f54888c38385d23228e80c58aa95e51ac814d8176bc441a1ea7c9e571839c

Observation 1aebbcfd-b34e-4ecb-a785-0269cc8888b5 · inbound

PGN: Design and Implementation of a Vision-Language Navigation System Based on Pangu Multimodal Foundation Model cites this paper.

PGN: Design and Implementation of a Vision-Language Navigation System Based on Pangu Multimodal Foundation Model NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:36:48.962835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:36:48.962835Z digest=sha256:283b44a2e64b8fe12448c544e094f074460c6b9773e21a6b09b5cef2df179833

Observation 87951153-cd2e-4ee6-a848-7ac3792bcd51 · inbound

ReferTrack: Referring Then Tracking for Embodied Visual Tracking cites this paper.

ReferTrack: Referring Then Tracking for Embodied Visual Tracking NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T11:00:24.420418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:00:24.420418Z digest=sha256:a938b2480a89a861420727ced6fe7746bf4e253b9731b64ef07fc8d4086b35ba