Pith. sign in

Paper Citation Record · LEDGER

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation

As of 12 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2605.22816.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.22816 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T04:46:03.020800Z

measured 56 of 56 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T00:30:53.972225Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

55 of 55 outbound references displayed

  • verified exact22
  • verified fuzzy32
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e6d2f198-dbe7-4194-adac-4ecadfcac18f · outbound

This paper cites Bevbert: Multimodal map pre-training for language-guided navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Bevbert: Multimodal map pre-training for language-guided navigation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.059909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:d95dae316cdcbc8942fcd40276fd0778e7902eeff35d97d5a015e68108d507c1

Observation 8186c9bc-05a4-4f68-9503-6031a5808e65 · outbound

This paper cites 1st place so- lutions for rxr-habitat vision-and-language navigation com- petition (cvpr 2022).

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation 1st place so- lutions for rxr-habitat vision-and-language navigation com- petition (cvpr 2022)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.113284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:94fe65dd2987eec89280c83390e4da85686f8b0f62a89a97198eeb76d7c5399b

Observation 699a5abb-f389-4291-b5f8-8be72a749eac · outbound

This paper cites Etpnav: Evolving topo- logical planning for vision-language navigation in continu- ous environments.IEEE Transactions on Pattern Analysis and Machine Intelligence.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Etpnav: Evolving topo- logical planning for vision-language navigation in continu- ous environments.IEEE Transactions on Pattern Analysis and Machine Intelligence

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.063270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:ad86abbe1c6fd6fb9bcb25af95067509686e994519d0beecd28107827ae6d503

Observation ebb697a6-b348-40dc-9333-05440049e02b · outbound

This paper cites On Evaluation of Embodied Navigation Agents.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation On Evaluation of Embodied Navigation Agents

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.584538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:17c0fd348ed2f7a372bce3e7ec1b72c22c4493b4743b70c4db2049cb7280e781

Observation dda6a7e9-9343-4262-aca6-5b94af8b6a42 · outbound

This paper cites Vision-and-language navigation: In- terpreting visually-grounded navigation instructions in real environments.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Vision-and-language navigation: In- terpreting visually-grounded navigation instructions in real environments

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.067491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:57e952c8ec92e7196220f1cac895007e3cd72f7b1e19e0b4894f44c632f56829

Observation edb0deb1-e9ee-4cb6-9e2e-27060b75b830 · outbound

This paper cites RT-H: Action Hierarchies Using Language.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation RT-H: Action Hierarchies Using Language

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.537564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:dd743c6942d0a22aac0c29af93fc3e81f2632b7b9089a452562faeba9c29663f

Observation 4505b0f1-72c5-4f90-91a8-ecdbc036c0e2 · outbound

This paper cites Matterport3d: Learning from rgb-d data in indoor environments.3DV.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Matterport3d: Learning from rgb-d data in indoor environments.3DV

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.128395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:6240bc63b5092c620eed8fdb75b2b28ad91d2dc5a78f214ed6a05a2a7805d43a

Observation 892cc9c5-a117-4872-9a48-a939f5fecce5 · outbound

This paper cites Object goal naviga- tion using goal-oriented semantic exploration.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Object goal naviga- tion using goal-oriented semantic exploration

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.109647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:5c62135cea96070818bb040662e3eb624cebef4029d71bc63396f107c27340e1

Observation d7205c7c-6876-465a-ae9a-310386ef4c43 · outbound

This paper cites Affordances-oriented planning using foundation models for continuous vision- language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Affordances-oriented planning using foundation models for continuous vision- language navigation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.157283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:ebe2f10994b274fdbff0a6dff008ba976b76280a3b45c7b467d425e6689c58c0

Observation eea711d6-b72c-406c-95fb-d30c8e7ece70 · outbound

This paper cites Topological planning with transform- ers for vision-and-language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Topological planning with transform- ers for vision-and-language navigation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.160530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:e3726592469e6db10d92e3367c18a246bd85ca86fb1f69f6bfb220ad485cd4f0

Observation 6f560508-131f-42c6-a24d-170b3f9b5cae · outbound

This paper cites Weakly-Supervised Multi-Granularity Map Learning for Vision-and-Language Navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Weakly-Supervised Multi-Granularity Map Learning for Vision-and-Language Navigation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.495232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:bd96bd3bd259d73aad372ce23ca55f00453ec175b0689cb2871fd50d5d0cff79

Observation a7d11f88-d4fd-4bce-93c9-5fddcb0893e1 · outbound

This paper cites $A^2$Nav: Action-Aware Zero-Shot Robot Navigation by Exploiting Vision-and-Language Ability of Foundation Models.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation $A^2$Nav: Action-Aware Zero-Shot Robot Navigation by Exploiting Vision-and-Language Ability of Foundation Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.595494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:dfa1c69ad7ab238022192a7e89eba035f99f7fabe4af302e138f5e4d7b7e2671

Observation 2120b067-32a6-4a06-8585-b872eaeb6435 · outbound

This paper cites NaVILA: Legged Robot Vision-Language-Action Model for Navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.562531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:aec48c2593ac0e1acc448698c19eede61f02d490cc9ef3eaa77b16f88deb5601

Observation 41058075-bcbf-4a78-9c25-1dada1c18bc0 · outbound

This paper cites Learning universal policies via text-guided video genera- tion.Advances in neural information processing systems, 36:9156–9172.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Learning universal policies via text-guided video genera- tion.Advances in neural information processing systems, 36:9156–9172

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.144290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:62685b2502fabdd61e4c25e3fa05bd67514d85a14071377e7a43dcfc6ceb77a1

Observation c76a6c9c-d031-4371-b704-dc4162fdc671 · outbound

This paper cites FLIP: Flow-Centric Generative Planning as General-Purpose Manipulation World Model.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation FLIP: Flow-Centric Generative Planning as General-Purpose Manipulation World Model

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.489225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:514ad397b3db2ca2efc370b7d1cfe4154712756cbf75119895c1577b9142a02e

Observation 7c76c7cd-a72b-4c66-a072-775f977ba613 · outbound

This paper cites OctoNav: Towards Generalist Embodied Navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation OctoNav: Towards Generalist Embodied Navigation

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.580136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:a2d02137fd0c9a95e32c1b034d01b296c9a00e0d691cbd4bfe2dcca2d406f56d

Observation d7a4b6d9-4f86-46a6-bb5b-8b5e954b8d57 · outbound

This paper cites VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.519797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:6e8719ad4c0e2353a4d2be78704bd4f532a25b8262282cfa06a4c1a887e3432e

Observation 0809e1ba-4f10-4929-8a6d-e002a2a6990d · outbound

This paper cites Cross-modal map learning for vision and language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Cross-modal map learning for vision and language navigation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.125774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:f6fd4b84a7bad8a42ae476107f4b720adc88fe5c2948902aa57dc39dc99cbfca

Observation 581958ed-c548-48f7-8cee-3216a5ea68a0 · outbound

This paper cites Bridg- ing the gap between learning in discrete and continuous en- vironments for vision-and-language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Bridg- ing the gap between learning in discrete and continuous en- vironments for vision-and-language navigation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.071811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:f3f2d64ba17b3e5b0d662670ae89140300ad466192b6d1faa9b3e7a7e1cb1272

Observation 947f2da4-d25b-42bb-8418-0b42aefce8d4 · outbound

This paper cites Learning navigational visual representations with semantic map su- pervision.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Learning navigational visual representations with semantic map su- pervision

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.082265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:d8ba0842ce9e478b489100dd84f100fb2ded93da89a1715c33b6b9da597e2948

Observation af2250b1-ef25-45e8-ad65-c7c28121e99e · outbound

This paper cites General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.501339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:5d5aec9538bac97f5e9df43a611a3dd2644c5c8b2e17fdb72854d2c3cff1b9f7

Observation a3465eff-ed0c-43bf-952c-9130d8632ef7 · outbound

This paper cites Robobrain: A unified brain model for robotic manipulation from abstract to concrete.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Robobrain: A unified brain model for robotic manipulation from abstract to concrete

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.106551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:ef0268aad8db600898fab5caf68c2053d37b316a1c15e5444d6c0d286d8dfd8e

Observation 9d408783-900b-4827-a7d7-be404d7e0699 · outbound

This paper cites Sim-2-sim transfer for vision- and-language navigation in continuous environments.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Sim-2-sim transfer for vision- and-language navigation in continuous environments

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.052944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:a388c8d0ca68ae166e830deb3debd475f95fe79d0f45ad16bc6343aa95ca1868

Observation 09fcebb2-f0bd-4ed5-a18a-defc949d6098 · outbound

This paper cites Beyond the nav-graph: Vision and language navigation in continuous environments.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Beyond the nav-graph: Vision and language navigation in continuous environments

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.056625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:ba52dd5a3680a16082dd2a170ef24dc43d59c42396b8bec1d6d8903af3607796

Observation 56229615-dd55-41dd-aba0-a73598ed21c2 · outbound

This paper cites Waypoint models for instruction- guided navigation in continuous environments.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Waypoint models for instruction- guided navigation in continuous environments

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.075633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:b79980be996d6a886cbcdb70e260b01e9979894d3f2cfdab2ca26879ed2c3b4d

Observation 11c05d4f-626b-4af1-ba7f-57be99b87fc2 · outbound

This paper cites Room-across-room: Multilingual vision- and-language navigation with dense spatiotemporal ground- ing.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Room-across-room: Multilingual vision- and-language navigation with dense spatiotemporal ground- ing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.140847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:6fbd780836ff1307f2df986248579864d4ea7586df5a138e27c320913c412b85

Observation 302351a1-f47c-48d5-8938-523d458b0480 · outbound

This paper cites Onetwovla: A unified vision-language-action model with adaptive reasoning.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Onetwovla: A unified vision-language-action model with adaptive reasoning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.513621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:497a0927a62a9e4a706e3c293847acd1ead7345f6b915be8b1ec8cca43ff2e08

Observation fda16bb8-09bf-4837-98ef-0e06788b4967 · outbound

This paper cites Nav-R1: Reasoning and Navigation in Embodied Scenes.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Nav-R1: Reasoning and Navigation in Embodied Scenes

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.556035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:98c7573ec16ff4db8333f9cd5171357eb6ed979d9fc33e874c698c84ba20f20f

Observation c77b0a6a-cb6c-4ab2-a815-44d30acfd687 · outbound

This paper cites Bird’s-eye-view scene graph for vision-language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Bird’s-eye-view scene graph for vision-language navigation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.116506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:97ae59fa21bb37ff4cfa4a8f43a5c30dfab4211c6237739575c7a363715cb6a5

Observation fb21af81-e66d-4377-b40d-b325c2858881 · outbound

This paper cites Dis- cuss before moving: Visual language navigation via multi- expert discussions.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Dis- cuss before moving: Visual language navigation via multi- expert discussions

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.119646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:a8194163489a4dd0b0f9ba3c1614defa411be20e5dcda2cf93cdf45cea4443a8

Observation 8ea6b500-6628-44aa-9cde-19551c7b3b51 · outbound

This paper cites Rt-affordance: Affordances are versatile intermediate rep- resentations for robot manipulation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Rt-affordance: Affordances are versatile intermediate rep- resentations for robot manipulation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.122790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:f4fcca3597439810f5bc19046a3050ed9bfc911cafb06a829ff9eb9d9c6cc7db

Observation 0ce6c4d5-b991-4c1b-a393-9f1103944045 · outbound

This paper cites VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.568388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:323aa3b741971d1c8c0be4fcc487963dffe20618e4476f65f3b55c660627b061

Observation 665f839b-6641-4fc4-b172-17863bd52b04 · outbound

This paper cites Language-aligned waypoint (law) super- vision for vision-and-language navigation in continuous en- vironments.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Language-aligned waypoint (law) super- vision for vision-and-language navigation in continuous en- vironments

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.099902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:67a6b6401589056554d946d12926c864cb30ecc21aef7b50a19d8c647ffbf597

Observation 685b7205-5126-4cb6-b2cb-1f1abe86f475 · outbound

This paper cites Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.600695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:f2f7302c85f1c37fc42151adb76aba38f2a328698b288144a3bb31ae9def6964

Observation 928f8690-1ae3-4010-ae78-dd6d9ac3b8cb · outbound

This paper cites Habitat: A platform for embodied ai research.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Habitat: A platform for embodied ai research

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.137637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:86a66e25540a7d1be3643bb56727019dca13fc233237672839a9476951c36047

Observation 68737d5f-53ef-4b44-a07f-3738098b9817 · outbound

This paper cites Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.531661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:050ede419578bc074b88d815cdd521d7cdad99e8e11676b17e52c2cbcdcb502a

Observation eece47ac-4fff-45f8-8a78-72074edc6ee4 · outbound

This paper cites Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:38:25.165602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:8136007031ab254b65faf6f7cdb1a77de8a02bb43c1d4fc9f056af8332759085

Observation 40a22603-89fd-4471-9d87-286a1c22c38b · outbound

This paper cites Dreamwalker: Mental planning for contin- uous vision-language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Dreamwalker: Mental planning for contin- uous vision-language navigation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.092805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:fc7081ec8cec16d09e069e93f18b8ed12e4781a22ad1c2b2bfc26900e6765e85

Observation 20316ea6-0082-4924-8bef-599d06814739 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.573900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:56f053cc010d715f73413752a2c02c96bd5660b24a3b8cedfc12ddf948987866

Observation d57f819d-65b9-4245-9f60-547f5136b448 · outbound

This paper cites Scaling data generation in vision-and-language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Scaling data generation in vision-and-language navigation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.150586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:038eff4a7c18643d4ae2d2624d74c768146e533f588e3ba125ef37abad9d135b

Observation 5c8b67ff-6f7c-4040-9a5b-48b0cf421cfa · outbound

This paper cites Gridmm: Grid memory map for vision- and-language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Gridmm: Grid memory map for vision- and-language navigation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.153491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:125632fbf38aeed2217226f34aa9a9071a7bed1af576ed75b2a54b4c8e931a71

Observation 5ed049f4-8bde-4fec-903d-b4d81db0e2e1 · outbound

This paper cites Lookahead exploration with neural radiance representation for continuous vision- language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Lookahead exploration with neural radiance representation for continuous vision- language navigation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.147514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:86d0b277e4deea8a768ac6fe3814f84dab504fa48a2f78aeae598fbec58c177a

Observation 55e1eabe-9bb0-488f-9ab8-ca85230911f0 · outbound

This paper cites Lookahead exploration with neural radiance representation for continuous vision- language navigation.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Lookahead exploration with neural radiance representation for continuous vision- language navigation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.089672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:503b36ed10581adf30fb44679fef12a61be7a5fa5f3bc7b9cfeee247fb129e39

Observation 1bbc662c-9436-4678-b822-7f7926c4fcc6 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models.Advances in neural information processing systems, 35:24824–24837.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Chain-of-thought prompting elicits reasoning in large lan- guage models.Advances in neural information processing systems, 35:24824–24837

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.096534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:a3e424e5a3211b02afbdd30f4133d8badf3a3d58a756d722abed46b1d45c3849

Observation 134d7494-b0f0-4396-bf3e-498cb0861b5b · outbound

This paper cites StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:4a1c709d0cda2dad3f74f5c7ac543e52149dc1c803391c3833c75afdee851761

Observation fcc3dab5-a1c7-4a4e-b1f7-4001c3a5b830 · outbound

This paper cites Any-point Trajectory Modeling for Policy Learning.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Any-point Trajectory Modeling for Policy Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.525691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:5d0259645b3792348839beb13373180d6ce479bcdd4ae3ce4800ae757cca3f53

Observation d8639eb4-61cb-46ae-8b52-22f21f54ccb1 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.549209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:44b93423cd471997eb7621e595d76402b68091bf5fa334f06fb78dd5c3193639

Observation 02b0d59f-7a48-462a-970e-22c96b6d04dd · outbound

This paper cites Learning Interactive Real-World Simulators.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Learning Interactive Real-World Simulators

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.481001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:c3630ac6e177b48c8e3392ffd0418f668a7020bb7cef4841dc9b94c0c2d17f8d

Observation 75e775aa-3980-439d-b79e-0aa6eec4ed1a · outbound

This paper cites CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:46:04.466276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:c8c934eb69c16e7e8c21ba5ae24981a53189b79e587d655e547478d94614f840

Observation c55ed96a-d93e-47a0-a7b9-e8b90a1349ca · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:46:04.590180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:d949663c7c38392fce710e3c513380ffbf7837263f77dec84bd061cfb4f3da2e

Observation 9676a36a-cd06-4854-bbfd-9a2bd7defbc7 · outbound

This paper cites Uni-navid: A video-based vision- language-action model for unifying embodied navigation tasks.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Uni-navid: A video-based vision- language-action model for unifying embodied navigation tasks

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.103248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:253f5e9ef5e9c22e0980acf74b2cd0559a4180080bd7ede1b04331341183b982

Observation 87cf9181-3a33-4f2a-aa9c-1032c9697bfb · outbound

This paper cites Navid: Video-based vlm plans the next step for vision-and-language navigation.Robotics: Science and Sys- tems.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Navid: Video-based vlm plans the next step for vision-and-language navigation.Robotics: Science and Sys- tems

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.131199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:be2dcaf630241e8bad06001794e9680d2aaedb1c01d3becd082b7b5ba1c5a02f

Observation fd420a47-33ad-4d6e-bbf0-c257ace5e15a · outbound

This paper cites Cot-vla: Visual chain-of-thought rea- soning for vision-language-action models.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Cot-vla: Visual chain-of-thought rea- soning for vision-language-action models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.134631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:f92987e884f28ddb7caab85271ca192c154292fa1f1879ddf732e98178217d8e

Observation dc2f9475-1544-41f5-850f-de4fccb0499f · outbound

This paper cites I see xxx. And I'm in xxx.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation I see xxx. And I'm in xxx

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:46:05.079023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:a9d56979e09ef01bff89fa03774c0642acbb2006fb44ab33013106ed0ca48091

Observation dcacd5a0-0491-4dc8-9ecd-9f7f26fd6e52 · outbound

This paper cites Our AwareVLN also achieves leading performance under this transfer setting, demonstrating strong robustness.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation Our AwareVLN also achieves leading performance under this transfer setting, demonstrating strong robustness

Reference 55

Resolution
malformed identifier
raw_fallback, observed 2026-05-22T04:46:05.085799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:63d2efe3a1a70ff2e5ffb57c9cc7f49ae56e5c51c29c55db86ed11c39f2d392d

Pith citing papers

Observation dfc96182-3778-4438-80e0-6e96e4aaae77 · inbound

CosFly-VLA: A Spatially Aware Vision-Language-Action Model for UAV Tracking cites this paper.

CosFly-VLA: A Spatially Aware Vision-Language-Action Model for UAV Tracking AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T00:30:53.972225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:30:53.972225Z digest=sha256:fd90ec8eab38ed600590963c0ec407d31e1c5abc626b1c8120e89e362b21cb5c