Pith. sign in

Paper Citation Record · LEDGER

Nav-R1: Reasoning and Navigation in Embodied Scenes

As of 21 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 11 inbound Pith citation observations for arXiv:2509.10884.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.10884 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T17:31:43.099505Z

measured 69 of 69 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:37:56.867359Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:45.302935Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved58
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9172adfc-2d8a-43c3-a5f0-634b92cc8ef4 · outbound

This paper cites Etpnav: Evolving topological planning for vision-language navigation in continuous environments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Etpnav: Evolving topological planning for vision-language navigation in continuous environments,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.015699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.015699Z digest=sha256:305858fb7789fc50f6d9ddff41c6a759480f70efc3c419e998de6a9463d504d4

Observation 2e9c351c-31e2-4a95-9f93-d2f390673c06 · outbound

This paper cites 1st place solutions for rxr-habitat vision-and-language navigation competition,.

Nav-R1: Reasoning and Navigation in Embodied Scenes 1st place solutions for rxr-habitat vision-and-language navigation competition,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.097078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.097078Z digest=sha256:c3595cf9631a79ef13990e7130c816d6d49800dd8fd33b9027de5383379a8d01

Observation e5f9e844-943e-4711-8b47-28ccad9e35e8 · outbound

This paper cites On Evaluation of Embodied Navigation Agents.

Nav-R1: Reasoning and Navigation in Embodied Scenes On Evaluation of Embodied Navigation Agents

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.233100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.233100Z digest=sha256:1b91881deef08a32877f79cb058511085894b8d1ffa6f3eef17a9476da95c0bf

Observation cc91307f-a7e1-4428-8e9c-da18fe07061b · outbound

This paper cites Vision-and-language nav- igation: Interpreting visually-grounded navigation instructions in real environments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Vision-and-language nav- igation: Interpreting visually-grounded navigation instructions in real environments,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.288992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.288992Z digest=sha256:32f3d52c4a72da9c1f5fd48c5c0ea32113fd8864f4bd504842b57cf21bd69124

Observation 23ff8f06-21f2-4cd1-8235-fafc0c84e08a · outbound

This paper cites METEOR: An automatic metric for MT evaluation with improved correlation with human judgments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes METEOR: An automatic metric for MT evaluation with improved correlation with human judgments,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.361054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.361054Z digest=sha256:2a7aaa5aacd6e8552fec232a4132f9153400b21e53fab791eb0ae930f1326df0

Observation 6e782dc0-a233-4efa-8b64-a8f9a6ff0242 · outbound

This paper cites Touchdown: Natural language navigation and spatial reasoning in visual street environments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Touchdown: Natural language navigation and spatial reasoning in visual street environments,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.436966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.436966Z digest=sha256:81a6609dc9f7eb5d0599c1755cd6106b380688e993924e76c7eeee54be69727a

Observation 3316db8f-2fdf-471a-84bb-97e85e6cf18e · outbound

This paper cites Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation.

Nav-R1: Reasoning and Navigation in Embodied Scenes Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.519836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.519836Z digest=sha256:be4999b1335678006ec6139475a2c1a26cd070ac3d4ad98f258da79b3dfce7fa

Observation 2075cca1-eccc-4eb4-add4-fdf0a55ccfef · outbound

This paper cites Topological planning with transformers for vision-and-language nav- igation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Topological planning with transformers for vision-and-language nav- igation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.581685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.581685Z digest=sha256:f7d4011573fd5c887e05bf0acdefd57d2253c3ca3519fcf9f6423c7521ec5a76

Observation f0c570ea-291a-45ab-bc56-91c86b6ebf15 · outbound

This paper cites Weakly- supervised multi-granularity map learning for vision-and-language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Weakly- supervised multi-granularity map learning for vision-and-language navigation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.637577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.637577Z digest=sha256:c462d704bad15cd66bb3133db13955a6753bf71a37a039b1f9b6eebbc4445bdd

Observation 43b2aefd-a524-479f-9283-2a0833a6ea5d · outbound

This paper cites Ll3da: Visual interactive instruction tuning for omni-3d understanding, reasoning, and planning,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Ll3da: Visual interactive instruction tuning for omni-3d understanding, reasoning, and planning,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.725346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.725346Z digest=sha256:00ca17acc10d5d4ced4b549b149b20ff21baa89dd5ea7ffee48a2f7edfe0b0f8

Observation 6eb69b24-e8c1-4332-bb3e-2b5f293abe54 · outbound

This paper cites NaVILA: Legged Robot Vision-Language-Action Model for Navigation.

Nav-R1: Reasoning and Navigation in Embodied Scenes NaVILA: Legged Robot Vision-Language-Action Model for Navigation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.800522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.800522Z digest=sha256:a027ca6b421a81f63f08dcf4812fe7f499f970ef733c6ddcfaf2e5f0834af407

Observation 42acfbc0-67f8-47c0-bd2e-d0fe3db4194e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Nav-R1: Reasoning and Navigation in Embodied Scenes DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.888941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.888941Z digest=sha256:b10db25cabd84115bb9f0c16d4f84d40ff8c40933794681a2a1100ae47f0eb49

Observation b78c5329-05cb-4dbb-a73c-462fdf91d666 · outbound

This paper cites OctoNav: Towards Generalist Embodied Navigation.

Nav-R1: Reasoning and Navigation in Embodied Scenes OctoNav: Towards Generalist Embodied Navigation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:38.957862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:38.957862Z digest=sha256:bf1c34292c725d49be4294f1d176a5bfcee46b4931fc1fc9503a72bde4670a4e

Observation 67359b30-e546-4933-ab67-470db9617423 · outbound

This paper cites Cross-modal map learning for vision and language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Cross-modal map learning for vision and language navigation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.022759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.022759Z digest=sha256:a2697d5bfc2691e7f3c8178f84a3443479ad23837e4af561387a8812dea9420d

Observation 941fcaf6-8663-445f-965e-9853f84b77fd · outbound

This paper cites GaussianVLM: Scene-centric 3D Vision-Language Models using Language-aligned Gaussian Splats for Embodied Reasoning and Beyond.

Nav-R1: Reasoning and Navigation in Embodied Scenes GaussianVLM: Scene-centric 3D Vision-Language Models using Language-aligned Gaussian Splats for Embodied Reasoning and Beyond

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.119379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.119379Z digest=sha256:60402b6478c66a3aed122eb5f606022608590804c5abdcb7ee817c62452d0745

Observation 0dc4115e-6e5c-4817-a8cd-0453ef5cc48b · outbound

This paper cites Bridging the gap between learning in discrete and continuous environments for vision-and- language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Bridging the gap between learning in discrete and continuous environments for vision-and- language navigation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.319240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.319240Z digest=sha256:8f50fdcb4a6848790e80d492aeecc534e472bacf17a69464b971b8024f5ce7f3

Observation 65142323-e9d2-4686-9f15-1d8002bb251b · outbound

This paper cites Learning navigational visual representations with semantic map supervision,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Learning navigational visual representations with semantic map supervision,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.454192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.454192Z digest=sha256:08359002aa9f130247af23234c024bd76881dedba6c181af5a7d456e36bfc196

Observation 3b8f2fd8-4ada-4a09-90a3-1ec9417531ed · outbound

This paper cites 3d- LLM: Injecting the 3d world into large language models,.

Nav-R1: Reasoning and Navigation in Embodied Scenes 3d- LLM: Injecting the 3d world into large language models,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.578629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.578629Z digest=sha256:f4feab6f5341685ebf64a812b8ea17537a898388b4a5966cba61e897e3ea11d5

Observation 8684212a-62de-4537-aaba-81cb7432693d · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

Nav-R1: Reasoning and Navigation in Embodied Scenes LoRA: Low-rank adaptation of large language models,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.760547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.760547Z digest=sha256:5f3e688ce5d54a87b238e3397e7dd0bddf6f67223281881240e1276b688123bc

Observation 063c6902-dd40-43bc-97f8-58b43b96ae69 · outbound

This paper cites Chat-scene: Bridging 3d scene and large language models with object identifiers,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Chat-scene: Bridging 3d scene and large language models with object identifiers,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:39.972130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:39.972130Z digest=sha256:dd9b5361d95020a09429e668c4619fbbd62319696092754b4cbdc47f02558665

Observation 8a76b0c3-461a-4806-a01c-60d58674bb28 · outbound

This paper cites An embodied generalist agent in 3d world,.

Nav-R1: Reasoning and Navigation in Embodied Scenes An embodied generalist agent in 3d world,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:40.120107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:40.120107Z digest=sha256:df45da039645363e86165b15eaf3925a6ead7b9c696bc770261030cf79f01e2d

Observation 07d64188-5cf4-46a2-8aee-3c300e39b994 · outbound

This paper cites An embodied generalist agent in 3d world,.

Nav-R1: Reasoning and Navigation in Embodied Scenes An embodied generalist agent in 3d world,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:40.255563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:40.255563Z digest=sha256:3863d175865139fb7e73609ae85193d938f2065952622430a7a7ec6e75e92995

Observation 3956be89-547f-4c8a-b341-c7ac0b8d38ce · outbound

This paper cites 3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding.

Nav-R1: Reasoning and Navigation in Embodied Scenes 3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:40.400707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:40.400707Z digest=sha256:b1d7ef884dd3a81a5bde24a94101dc99c293b123880d4e01c3da6ebf667b62a5

Observation 23ae5b24-bd45-432a-82b0-285805b76064 · outbound

This paper cites 3D CoCa: Contrastive Learners are 3D Captioners.

Nav-R1: Reasoning and Navigation in Embodied Scenes 3D CoCa: Contrastive Learners are 3D Captioners

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:40.570298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:40.570298Z digest=sha256:e2562e4cdcab3feb2a21209e2fd9979a4a7add60501a9963a2e6ee772bcabf44

Observation 1b58e7ea-5237-4524-a269-f9889b875a63 · outbound

This paper cites DC-Scene: Data-Centric Learning for 3D Scene Understanding.

Nav-R1: Reasoning and Navigation in Embodied Scenes DC-Scene: Data-Centric Learning for 3D Scene Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:40.726038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:40.726038Z digest=sha256:3c49b4c6312749f8fe21dc703763aa64aacd2b8ff50560c8eeb8e0236b70ff0c

Observation 15de6308-2fec-4108-913a-09e0e13603e9 · outbound

This paper cites General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping.

Nav-R1: Reasoning and Navigation in Embodied Scenes General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:40.857383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:40.857383Z digest=sha256:488af0f10b9641518aba996281a60d46478bf67ae159a149484419e4d6b63a24

Observation 9de62c38-ad49-43ae-aed4-6b9f364d512f · outbound

This paper cites Kahneman,Thinking, Fast and Slow.

Nav-R1: Reasoning and Navigation in Embodied Scenes Kahneman,Thinking, Fast and Slow

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.009140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.009140Z digest=sha256:6822393303438ab7ad6750c36a737e01b817383c5602db92cadc61d286078df5

Observation dcecca26-3e50-460a-9031-3c61370cb6f4 · outbound

This paper cites Sim-2-sim transfer for vision-and-language navigation in continuous environments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Sim-2-sim transfer for vision-and-language navigation in continuous environments,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.151069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.151069Z digest=sha256:07404381225a228ecff96c29f0b4644960e5ec6e699f8e6dd003ff39fba43134

Observation 5facd0a3-2c14-45a7-b728-c1a7a37da7ad · outbound

This paper cites Beyond the nav-graph: Vision-and-language navigation in continuous environ- ments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Beyond the nav-graph: Vision-and-language navigation in continuous environ- ments,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.289147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.289147Z digest=sha256:b0203be15f21b7012c8fd92aaa68bc16b72f170facd4eda2ce3bcdf52cbbaff2

Observation 5c492863-6ad6-4676-937f-c868caf15d54 · outbound

This paper cites Room- across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Room- across-room: Multilingual vision-and-language navigation with dense spatiotemporal grounding,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.413264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.413264Z digest=sha256:1baf0a28ec6f90f9f12637b8ead942d62352fd3baa3a477b3783f28e2c6bc345

Observation f07aa857-fecb-4ac9-9359-0008ce98d4da · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries,.

Nav-R1: Reasoning and Navigation in Embodied Scenes ROUGE: A package for automatic evaluation of summaries,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.509811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.509811Z digest=sha256:75e04c916e5fcd16bc8de6b4a3244f564033533f3900baedf339ba6bb2428663

Observation 82cb448a-2296-4dc1-9c56-378ad85f4af8 · outbound

This paper cites InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment.

Nav-R1: Reasoning and Navigation in Embodied Scenes InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.645023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.645023Z digest=sha256:ef07157f6ef3c7b9e1fb9ba394f63df5a77b016cc90820ae742a608c5184005b

Observation 2d2375a5-f519-44b8-af75-1ba13ac3de8b · outbound

This paper cites Sqa3d: Situated question answering in 3d scenes,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Sqa3d: Situated question answering in 3d scenes,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.767705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.767705Z digest=sha256:65e767c998bf83ffb1e40a90e9af5425f5cfdbf00f80aec09ed9c9cfb7c102c6

Observation 84864beb-e484-4b38-8b6b-743eb43bbe32 · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Bleu: a method for automatic evaluation of machine translation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:41.897746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:41.897746Z digest=sha256:752fdbb9580e9f046aedeba9848eae11cd6c81d7afe9de8399df04d01d9d203b

Observation 5a43ff08-9bbf-436f-9b16-31ef267dc817 · outbound

This paper cites Alvinn: an autonomous land vehicle in a neural network,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Alvinn: an autonomous land vehicle in a neural network,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.059993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.059993Z digest=sha256:63cc7af3e5cd64febc2a9c064c7a2104356a46de31765182b178937596a87e0b

Observation b85f0e97-f99f-4328-b6a6-3b51bc81b640 · outbound

This paper cites VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning.

Nav-R1: Reasoning and Navigation in Embodied Scenes VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.089676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.089676Z digest=sha256:b615b4c9dfe781ab56e808793f1161b91e753b0ee8841a1ab0700433df8e0fa7

Observation caf002e9-44d8-4df7-9433-efcd2ef8560e · outbound

This paper cites Language- aligned waypoint (law) supervision for vision-and-language navigation in continuous environments,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Language- aligned waypoint (law) supervision for vision-and-language navigation in continuous environments,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.223630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.223630Z digest=sha256:829aefcb30daad52350265a8e50bc1d466388539ad07dd87176b4e6c75bbab12

Observation a550668a-3a25-4dce-9394-98ccc23181a2 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning,.

Nav-R1: Reasoning and Navigation in Embodied Scenes A reduction of imitation learning and structured prediction to no-regret online learning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.316143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.316143Z digest=sha256:31b6807076b2de3a7fc48115958e1e676ba1eb8354e7d4477be43ec9737c568e

Observation 207ca791-c1dc-4615-ad0f-89b40581fd5d · outbound

This paper cites Proximal Policy Optimization Algorithms.

Nav-R1: Reasoning and Navigation in Embodied Scenes Proximal Policy Optimization Algorithms

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.321667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.321667Z digest=sha256:db6f62ba7d0979e73302fd6a84db8806692691dc89e6ca72062005f04b3e20bf

Observation 5ac44db6-3983-4df0-abac-761969d547e0 · outbound

This paper cites Hazards in daily life? enabling robots to proactively detect and resolve anomalies,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Hazards in daily life? enabling robots to proactively detect and resolve anomalies,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.415632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.415632Z digest=sha256:5e0f26288fc127338ad48bc7565c62547c6327bad6756436e982a013f938dcde

Observation fdb7487d-55ca-4d5f-8c8d-b70dd774a0d9 · outbound

This paper cites ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models.

Nav-R1: Reasoning and Navigation in Embodied Scenes ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.500658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.500658Z digest=sha256:b58ddded8eb400469436be8b9087dbf33a5eac00708e927d60eb652815372bd2

Observation f4842735-bac2-4fc9-a72d-259c01aeecf8 · outbound

This paper cites Evaluating Gemini in an arena for learning.

Nav-R1: Reasoning and Navigation in Embodied Scenes Evaluating Gemini in an arena for learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.588126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.588126Z digest=sha256:209a259e85023d031ca0b4362eb3a0e2d5012a667d5d840f6e198c832057f519

Observation ab7e0219-8c7e-4991-aa4a-d2d180a7e1e4 · outbound

This paper cites Cider: Consensus-based image description evaluation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Cider: Consensus-based image description evaluation,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.648263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.648263Z digest=sha256:3021491190ddfdcc1a3b6d82ac014d635ab21d5caf135a62b9426c321755de86

Observation 51c4d7bf-a5af-40ac-b5ad-eac161090364 · outbound

This paper cites Dreamwalker: Mental planning for continuous vision-language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Dreamwalker: Mental planning for continuous vision-language navigation,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.686756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.686756Z digest=sha256:a1f7a77c6190b0366de4b5fdc54970e29883a70cb2a764dde573e76f2b6e23bf

Observation 80d294b2-9c95-457c-b096-5ae05da9346e · outbound

This paper cites Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models.

Nav-R1: Reasoning and Navigation in Embodied Scenes Spatial 3D-LLM: Exploring Spatial Awareness in 3D Vision-Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.766632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.766632Z digest=sha256:526866c5764693f46267adc30f89f821eed3d513cbf15f4be253e9c066e9ce52

Observation af75a56a-cc84-4fa9-8f6e-7e740b2b5be8 · outbound

This paper cites Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.821024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.821024Z digest=sha256:0e93b1d3aeaf2e329f9b4ef9820fcc715de7aaef1cc38af3a64eadc8976c9042

Observation de4913cd-8648-42dd-9cb2-4802705fe941 · outbound

This paper cites Looka- head exploration with neural radiance representation for continuous vision-language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Looka- head exploration with neural radiance representation for continuous vision-language navigation,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.875262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.875262Z digest=sha256:6416c69dae04441e3ff174a4326dc8d6668c077a8da9494669ba02e225980dd5

Observation 4f41e7c9-efb9-49fa-8194-50ee39ceb4c5 · outbound

This paper cites Gridmm: Grid memory map for vision-and-language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Gridmm: Grid memory map for vision-and-language navigation,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.931648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.931648Z digest=sha256:9a15fdb8e9c8e19b41341a83e7d819506c9ef7ce07880c37ac71ec2694984fe9

Observation 4587285b-ab9f-4591-ad7c-fed4f4c6fc10 · outbound

This paper cites StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling.

Nav-R1: Reasoning and Navigation in Embodied Scenes StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.003448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.003448Z digest=sha256:a6b047f208a1aceb4cd5a82d48e1dff4d5cd25390202a7adeeae24a604ec2389

Observation d0766025-4193-4401-a796-2aeb7d174bd0 · outbound

This paper cites Vlfm: Vision- language frontier maps for zero-shot semantic navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Vlfm: Vision- language frontier maps for zero-shot semantic navigation,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.051782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.051782Z digest=sha256:bde65893b1e8c9d674ee1727a955b6f76692872531c3d764ea61a8fb9844ebe7

Observation 86b88114-8f1c-4211-96a2-adee934fc2b5 · outbound

This paper cites Hm3d- ovon: A dataset and benchmark for open-vocabulary object goal navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Hm3d- ovon: A dataset and benchmark for open-vocabulary object goal navigation,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.068835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.068835Z digest=sha256:c7876128aa979ad94bca68c1bd98fcdb71a632e3fe5d7d74e5a7fa40bfa7a778

Observation 18c9cfa3-eaa7-4c14-a779-f7bfc39d3d36 · outbound

This paper cites CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model.

Nav-R1: Reasoning and Navigation in Embodied Scenes CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.075074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.075074Z digest=sha256:3115394b02f86c110a651b8142dc1cff28753d67b3e9403e7bf3c4675ddb83ca

Observation 483c3ef8-5c46-46b4-9cdf-94715998569f · outbound

This paper cites Uni-navid: A video-based vision-language- action model for unifying embodied navigation tasks,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Uni-navid: A video-based vision-language- action model for unifying embodied navigation tasks,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.080070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.080070Z digest=sha256:05d20841780573f8f7a8f66e1986c5d4a463f8ef592677c79a0c82d0b96aba4a

Observation e4c176e0-1d5a-481b-93e5-fcec5e864b20 · outbound

This paper cites Navid: Video-based vlm plans the next step for vision-and-language navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Navid: Video-based vlm plans the next step for vision-and-language navigation,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.085265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.085265Z digest=sha256:d4ac7def526a475e1c9b8319067aa8686f39ba851978c8b98e87dbad5e8024ce

Observation c02b4ead-fde1-4b93-92b0-417ce4d8e435 · outbound

This paper cites LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences.

Nav-R1: Reasoning and Navigation in Embodied Scenes LSceneLLM: Enhancing Large 3D Scene Understanding Using Adaptive Visual Preferences

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.089996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.089996Z digest=sha256:f0610ab25fda1cd6d3fc365ae0f8a1c81539605b32fa0cabfeadc871ee058042

Observation b83fb881-773f-4c53-9d4f-2c5147686616 · outbound

This paper cites Soon: Scenario oriented object navigation with graph-based exploration,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Soon: Scenario oriented object navigation with graph-based exploration,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.094703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.094703Z digest=sha256:e1762658871df107ce2221622bfa3c5daa3be58fec030dbb987b701a9a506581

Observation bb929dca-8eb3-48cf-9b8c-ae16f8004149 · outbound

This paper cites Move to understand a 3d scene: Bridging visual grounding and exploration for efficient and versatile embodied navigation,.

Nav-R1: Reasoning and Navigation in Embodied Scenes Move to understand a 3d scene: Bridging visual grounding and exploration for efficient and versatile embodied navigation,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.099505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.099505Z digest=sha256:4c907ef09067806cd335834541b602a09626723ae43ea2e3e4a9951616997a32

Observation 2136eca0-145a-4c23-91ed-2fd2d091fc05 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Nav-R1: Reasoning and Navigation in Embodied Scenes DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:42.385017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:42.385017Z digest=sha256:5f5ac49db75f0bc8a999511c45106051247563bd61865fc79b7bdac0938a782e

Pith citing papers

Observation 97f9a599-f2df-4782-be89-9de0fe60b0d1 · inbound

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation cites this paper.

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:05:15.957964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T21:02:38.013115Z digest=sha256:04c661c84fc7ce84ce78689574c0098f2a0a3778a5965437e219816d22d3131f

Observation 88d9edd0-b521-4645-a5e2-a78cab1f5504 · inbound

GeoWorld: Geometric World Models cites this paper.

GeoWorld: Geometric World Models Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:40:03.318232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T11:39:15.308355Z digest=sha256:211cbaf49f37a5b4b8321b7dc6dea1286d9741714ba68a574a85126fcda0350e

Observation 49ed771a-63d4-44aa-869a-bf3199c9295d · inbound

What if? Emulative Simulation with World Models for Situated Reasoning cites this paper.

What if? Emulative Simulation with World Models for Situated Reasoning Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-15T13:51:30.008232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:51:30.008232Z digest=sha256:0e5929a9905e374c63ad0a10b4242a13963711192ca168a135878d1538c2cb6b

Observation 2e017691-05b0-4605-b195-6456ffd25768 · inbound

UniMesh: Unifying 3D Mesh Understanding and Generation cites this paper.

UniMesh: Unifying 3D Mesh Understanding and Generation Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:56:47.345040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T06:55:42.679323Z digest=sha256:b0046ecb2fb81947a9af24b615e203ee276b325bb46bef8177c9a1eb6c348a17

Observation 62431785-4ba0-4496-82ef-121c8f058ed2 · inbound

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation cites this paper.

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:31.239026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T05:05:33.606975Z digest=sha256:96ce3657ef7f021b75d4adfd4107a127e9e1df8672718d0ca187f69a2f08012e

Observation fda16bb8-09bf-4837-98ef-0e06788b4967 · inbound

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation cites this paper.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.556035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:ba43e3ca52165feacacdf6d775137ec0c12f5f312adda8312be03c5ca2c8f980

Observation 0554d21b-c375-44db-93b8-3ff12891f63a · inbound

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation cites this paper.

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:15.792458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T15:39:13.803571Z digest=sha256:6d441a0fcc7291836d053cf4f9181c475585f7fc0a83d735c4d149056fa9a3e6

Observation 27e5eb30-faad-40e2-a406-af8102b2f0b9 · inbound

PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological Maps cites this paper.

PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological Maps Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:16.292609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T15:35:47.290488Z digest=sha256:40734e0539b7f8598ab8f94ef4565fbd5a775cdb87161111dd54de4cda3d661e

Observation a20c37a9-863d-4104-b03c-1452e0c439c2 · inbound

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views cites this paper.

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:39:45.305524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T08:38:46.044079Z digest=sha256:5077852838e5e44ace7ff032c74794c2deb307f31e64df58a325064d6a708ff8

Observation 9a79b739-c8c8-4553-bd3a-b22fa893aa3d · inbound

ConsiSpace: Learning Geometric Consistency Matters for Video Spatial Reasoning cites this paper.

ConsiSpace: Learning Geometric Consistency Matters for Video Spatial Reasoning Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T17:36:38.963388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:36:38.963388Z digest=sha256:3b976b4796a16d1846f15d80d2fc143f095ac1d119c42aa12647fa09df174ade

Observation 44927f04-6f27-4de1-a4f8-115fa830b0d7 · inbound

From Failures to Supervision: DynamicEnvPlan for Robust Long-Horizon Embodied Planning cites this paper.

From Failures to Supervision: DynamicEnvPlan for Robust Long-Horizon Embodied Planning Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T00:37:56.867359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:37:56.867359Z digest=sha256:5546909e723de661d69fddf8f1221537ff8e581a4b815e41b1809aad46573e2b