Pith. sign in

Paper Citation Record · LEDGER

VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2506.17221.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17221 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T19:24:39.321499Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.489833Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8f6d4493-0bc1-4a99-9994-6333f0c359cd · inbound

IRPO: Boosting Image Restoration via Post-training GRPO cites this paper.

IRPO: Boosting Image Restoration via Post-training GRPO VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T19:24:39.321499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:24:39.321499Z digest=sha256:e7449866771bf3d936d3fb3def1ec0a2c30163ed1b778640699087d50fb66195

Observation 49bc9c5f-727f-4eb4-9bc8-98d0d5dd10a2 · inbound

Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning cites this paper.

Aerial Vision-Language Navigation with a Unified Framework for Spatial, Temporal and Embodied Reasoning VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:58:42.589185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-16T23:58:40.992942Z digest=sha256:0281ca34b42dd962cb087fdb61ae11807caee1416441761020c1dba9cfc574cf

Observation 6ef87bd5-bd44-4240-8069-7b85f724ef06 · inbound

Token Warping Helps MLLMs Look from Nearby Viewpoints cites this paper.

Token Warping Helps MLLMs Look from Nearby Viewpoints VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:08:17.442855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-13T21:07:55.062113Z digest=sha256:9341f591d1147bc88095164a33465460239e1d43a6d85a4b5cc1a7c8c107821a

Observation 818345e2-d07f-4a43-a906-1cca8202e0c9 · inbound

Hypothesis Graph Refinement: Hypothesis-Driven Exploration with Cascade Error Correction for Embodied Navigation cites this paper.

Hypothesis Graph Refinement: Hypothesis-Driven Exploration with Cascade Error Correction for Embodied Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:00.907219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-13T17:05:57.205365Z digest=sha256:103622b541b5f82a4d4fa040e9edf1cc5c999df7c800f7257024665c374f9068

Observation fde93708-7efc-41f8-aaa4-f929867df094 · inbound

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning cites this paper.

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:47.751455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T19:16:46.753641Z digest=sha256:cb25e7b42c44ad3fd65b8b5565804d3f4f448294319a461fa1ec26c2a39b0fab

Observation 53443418-d67e-4f2a-9708-e9072ca1faf8 · inbound

HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation cites this paper.

HiRO-Nav: Hybrid ReasOning Enables Efficient Embodied Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:41:01.531030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T17:01:13.350219Z digest=sha256:776e0fcb03ecd5b4d07ab6382d14dc749a24a3b5e728fd6fbf55178bb8fa8b78

Observation 062ae992-5bef-4e9a-af06-a426058fbc1f · inbound

Think before Go: Hierarchical Reasoning for Image-goal Navigation cites this paper.

Think before Go: Hierarchical Reasoning for Image-goal Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:10.485653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-10T05:43:27.972164Z digest=sha256:d6a809ce2ada287846ffb0980539e371802a8523ba2b5984c0114da32370a816

Observation d7e87bfa-b62d-4502-b272-c4e2f2ccce5b · inbound

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation cites this paper.

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:51:45.828270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-10T06:50:34.310831Z digest=sha256:516ab999aad3f9304950c2b00854a6b2891be63ac7c336af26108117e2c26a67

Observation 9f612723-deee-4adf-a40e-1a7f51c2744e · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 77

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.491147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:c996a68390da31339db17f8d05e2b3c669a94ed32cfc3215cb0c8350daeff96a

Observation c75ecfb5-fb62-42f1-a8d0-f4067bd19a30 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:10.296777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:5ce195f85f029487f0f5532f18016c54e29a37e77b56f8b71a6e1dd0d0001cda

Observation e37d385e-ef1d-407e-acb5-948dbc1a8cc6 · inbound

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation cites this paper.

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:36:31.243113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:05:33.606975Z digest=sha256:59b14d46e60e8e2e1cbb6c9c738f2c60dc6e3e6ed2e2dc1a93991bcb8a4788d0

Observation 79ce5e0b-fb83-47c5-bf81-7c5e0cf19292 · inbound

Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search cites this paper.

Beyond Thinking: Imagining in 360$^\circ$ for Humanoid Visual Search VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:01:18.332757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-05-12T02:58:46.728868Z digest=sha256:3bf196333b591a1400b5c53ad8940c53c99f74b48d748d54085c729f5029207d

Observation f66da04d-f31d-4ccc-9688-5ff8f158570e · inbound

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation cites this paper.

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-20T17:48:48.633274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-20T17:47:32.953903Z digest=sha256:713ea9f72dd706ed9c9ec373210f340039a78771cb3c2fbc64ca549e55169bd1

Observation 0ce6c4d5-b991-4c1b-a393-9f1103944045 · inbound

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation cites this paper.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.568388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:96c38f35cbf5032bdcd418c5d015482f45d517043c930aa8785cfebb1a17782f

Observation 8da6f638-a852-4621-a212-12dc15632643 · inbound

World Models as Group Actions cites this paper.

World Models as Group Actions VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.331396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T13:17:42.720825Z digest=sha256:38230fdfcbb0aa47cf4aa195ac53d0f3d06c2a3e990a8d5f30a5df6ffbf0d697

Observation ed9be44b-6172-4d41-b974-7b2b4d67a4ed · inbound

Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation cites this paper.

Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:34:01.336932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-29T22:33:32.671550Z digest=sha256:f0b500f6b7ca2541ff637b6907ec40e65c80efb308363c6db414a17e0fb4cb66

Observation b0a0df41-f434-4da9-90ef-c03458f72283 · inbound

Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning cites this paper.

Reasmory: 3D Reconstruction as Explicit Memory for VLMs Spatial Reasoning VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:46:13.683917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-28T17:46:21.821764Z digest=sha256:e6972041dea79c2f704b02cf8bf99f9beb98e02839d89b209201964399007f25

Observation 59072d98-3d15-47d9-b66f-d43b0980ef2a · inbound

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation cites this paper.

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:16:15.779796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-28T15:39:13.803571Z digest=sha256:8cb8ec42325a1c64d8b37045de0355ce51090492cf852e5ef832f36e38b83389

Observation 39828a1e-b0db-486a-b879-ad91aba9459b · inbound

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation cites this paper.

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.765948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T21:40:58.329546Z digest=sha256:ac4bacdc51d05e4bdeac39df53b91c61a4dc30b93c285b2978900ce01c737b82

Observation 9ee761ed-2ee7-4e27-8aab-274aa4eee27c · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 239

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.728048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:ba6a34dc7c5eccd1a9c1df7976bdacfe40ea6db2ffba6d64db8b0cf4bc9acde9

Observation afdb57f2-0228-4d53-b770-cbc3316ed17e · inbound

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation cites this paper.

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:08:21.390625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-07-03T14:04:46.400336Z digest=sha256:f21890c3a4b49ef25f9bde3809c7fdff6a81fdad416146f11e169b36c7e465d3

Observation 43227267-f85c-427b-b2e8-587b41454996 · inbound

A Comprehensive Survey and Systematic Real-World Evaluation of Embodied Vision-and-Language Navigation cites this paper.

A Comprehensive Survey and Systematic Real-World Evaluation of Embodied Vision-and-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T15:39:59.878169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:39:59.878169Z digest=sha256:ca4a32ef8eb6a9f1bb08aca57653991e6f953e665436da44b0cc0d2f1b805bf8

Observation 86424cd5-e548-4337-b63d-c8f01723c1ad · inbound

Joint On-and-Off Policy Learning for Vision-and-Language Navigation cites this paper.

Joint On-and-Off Policy Learning for Vision-and-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T05:13:54.793452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:13:54.793452Z digest=sha256:d1ed4199b07035e8b58a0d98de654d7a4ce0be39e41b5690e06629ccb99a4999

Observation 7b9be73c-fe2f-4bf4-b991-db29888e2848 · inbound

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation cites this paper.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.484911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.484911Z digest=sha256:b11f8155dcd3ff3a569ae8b6e85df8c48785c15d7ecd11440887fad0e6405696