Pith. sign in

Paper Citation Record · LEDGER

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation

As of 3 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2607.23504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23504 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-30T20:42:04.352249Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved47
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4249682e-8f12-4b2b-ae4f-320f69361e1c · outbound

This paper cites GPT-4 Technical Report.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.541961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.541961Z digest=sha256:fe0bead8bbbabe0639af6b439bc627c103913f131193c1f68f4f11e6420c2be0

Observation e67b5f55-f346-4833-8788-7def70da434e · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35: 23716–23736, 2022.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35: 23716–23736, 2022

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.628557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.628557Z digest=sha256:d9b6b6091b690544aed9143f42da5cbfd55f56e0021bbb007c912f88e5ddafd7

Observation c85c0990-00a8-4e6d-b5d4-ad78ce59a3fe · outbound

This paper cites Etpnav: Evolving topological planning for vision-language navigation in continuous environments.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Etpnav: Evolving topological planning for vision-language navigation in continuous environments.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.675416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.675416Z digest=sha256:0e57009a059601e04c7a6036376a97cc2b49d785b4f66f3a0ec8d0367b4b1795

Observation 166c0485-c004-449f-81b5-7ca6189cc13a · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.715588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.715588Z digest=sha256:d97b0864db6feae3b525a58f7ebfeac5f1dff16e9c0455d3cff284d3ce4a1af0

Observation c924cc41-2f60-424d-9c01-ac2f39d972f8 · outbound

This paper cites Sim-to-real transfer for vision-and-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Sim-to-real transfer for vision-and-language navigation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.785592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.785592Z digest=sha256:dfbb3933eb807006df97333d011785225e915bf5ea116adc58341973ca0be478

Observation 17d349f9-9cfd-4e5f-a6e6-14f848601dc3 · outbound

This paper cites Qwen3-VL Technical Report.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Qwen3-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.854773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.854773Z digest=sha256:56af6e15c0a1ca00b7c902c3146187c7e54679db3df7809582e4767f4815305a

Observation 338138f6-32f2-45ea-a9a7-c25effd74dff · outbound

This paper cites Topological planning with transformers for vision-and-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Topological planning with transformers for vision-and-language navigation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.903996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.903996Z digest=sha256:98be29bee5ada2c64691fee6e6c5ed409d2b7f746fe13c57693ff2026e9a7e1a

Observation 16cca270-3de1-432f-8d88-495bcdfd5430 · outbound

This paper cites Weakly-supervised multi-granularity map learning for vision-and-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Weakly-supervised multi-granularity map learning for vision-and-language navigation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:02.981520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:02.981520Z digest=sha256:b88d7ef7219f3d389e8a0f7407ea9935e3c97950d2f4e54c812b25036864e3f6

Observation 51a0af7a-2ead-4729-b06b-9357ef1cda00 · outbound

This paper cites NaVILA: Legged Robot Vision-Language-Action Model for Navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation NaVILA: Legged Robot Vision-Language-Action Model for Navigation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.068111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.068111Z digest=sha256:9551a3b0b28e873927ee78f7f81a2a360f58d9284b68c1680af11bf0d87c611d

Observation c86f0159-f8d6-4547-8650-1b13c692fc31 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.137249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.137249Z digest=sha256:a2067d03d8052044fa58a32f8b4a0eac8e5cdec3a8a17dc64c64b4bd3cfd6a1b

Observation 94c847d0-1e6b-40a3-a78f-550293eed3ad · outbound

This paper cites Speaker- follower models for vision-and-language navigation.Advances in neural information processing systems, 31, 2018.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Speaker- follower models for vision-and-language navigation.Advances in neural information processing systems, 31, 2018

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.197473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.197473Z digest=sha256:fb6779c174a9f6080cfe44899afc0e73ba6bea50cdec2efb3f4ff0a2c54afd67

Observation 4f368302-63c0-4c80-a8ba-595c46d7c586 · outbound

This paper cites Counterfactual vision-and-language navigation via adversarial path sampler.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Counterfactual vision-and-language navigation via adversarial path sampler

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.241149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.241149Z digest=sha256:84142e85a430ef1ef8aecd33dd77653b3e2bf2a19f47049623c6f38aab9e0c51

Observation e55bdd75-8713-4096-8559-266d6033c395 · outbound

This paper cites Cross-modal map learning for vision and language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Cross-modal map learning for vision and language navigation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.299565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.299565Z digest=sha256:c027cde41723b2ca07ac962ea3990b72f41a2af053feaabfb49edff5cffec154

Observation 270290de-cb87-4127-9f80-2746f625f5c3 · outbound

This paper cites Language and visual entity relationship graph for agent navigation.Advances in Neural Information Processing Systems, 33:7685–7696, 2020.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Language and visual entity relationship graph for agent navigation.Advances in Neural Information Processing Systems, 33:7685–7696, 2020

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.349403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.349403Z digest=sha256:c7a77279ecac6c25b9e4d307ddfdcd39153ecb84fa5fe2ac07a24eb06350dcbf

Observation a6b9bfce-89b2-4c02-971b-9be105d8bd8f · outbound

This paper cites Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.414738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.414738Z digest=sha256:a86dae1885f9c603a03edd75214f5c9270dc6693ecaf24f35d87b8273e53f798

Observation f11714af-b5db-4d05-8e1d-824a81b8b085 · outbound

This paper cites Sim-2-sim transfer for vision-and-language navigation in con- tinuous environments.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Sim-2-sim transfer for vision-and-language navigation in con- tinuous environments

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.492588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.492588Z digest=sha256:355334aaec8df987b7e63444ba13cd9404c5ec3da5d78ff5425c6674e104b11c

Observation 3ebf72e2-17f2-4326-92d3-cde805daa004 · outbound

This paper cites Beyond the nav- graph: Vision-and-language navigation in continuous environments.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Beyond the nav- graph: Vision-and-language navigation in continuous environments

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.547400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.547400Z digest=sha256:f68a48cc3acc2ff9a475c7155f6ca5e84bea85d13f1a7a60ff4d5a3b22168e32

Observation b69145a9-ff25-4ff7-a389-99f55bd69f92 · outbound

This paper cites Waypoint models for instruction-guided navigation in continuous environments.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Waypoint models for instruction-guided navigation in continuous environments

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.624402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.624402Z digest=sha256:1f8a00fb6efca199c6cec461be6f04bcd45a5634fb493b9b11e547a9d5b466e8

Observation 830de76f-17a1-4e7d-9bae-0b9a47022a18 · outbound

This paper cites Room-across- room: Multilingual vision-and-language navigation with dense spatiotemporal grounding.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Room-across- room: Multilingual vision-and-language navigation with dense spatiotemporal grounding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.706075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.706075Z digest=sha256:b44ae00b0c0b80bb3df23dbb79f36769b0fc7dc2d67d1e039d433fe439219cdc

Observation 77161361-f183-4dea-b951-ef5468de47e1 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.770114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.770114Z digest=sha256:add501e3e49d09c07fa7da5be1e4641c5ab384d7bf9f537a602b317bfdc1f647

Observation 05df4398-8db7-45a9-bdc6-e94aae7c0cc0 · outbound

This paper cites Llama-vid: An image is worth 2 tokens in large language models.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Llama-vid: An image is worth 2 tokens in large language models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.851911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.851911Z digest=sha256:62b14a61a4b4583627bfde84275896c7eebabd19ecdbf92ee12d86065548273b

Observation a1967f81-244f-4340-a806-6019a8046a72 · outbound

This paper cites Vila: On pre-training for visual language models.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Vila: On pre-training for visual language models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.945113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.945113Z digest=sha256:bc6bc2c6a210c5ffa92cdbe156d15c7c1b72807a690fd5d31707634d0c1167c4

Observation 6753bae8-a674-4081-8b2d-524570ad9582 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:03.969709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:03.969709Z digest=sha256:7494fb0492471167a94dd21f87e9cc176a4499a032eb30b7c9710eadf0d6f47d

Observation bf81f264-832a-4a0b-9f3e-efb10f36af84 · outbound

This paper cites InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.019111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.019111Z digest=sha256:655bf0f64891328aec3d60265db229e42cb040f7d53163e103cfbab89f801098

Observation ae4c0ceb-7a80-412a-87d0-b5379d9e0139 · outbound

This paper cites Discuss before moving: Visual language navigation via multi-expert discussions.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Discuss before moving: Visual language navigation via multi-expert discussions

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.124407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.124407Z digest=sha256:ca3eafbb1a43ad8a8d51bcef1089dd2789740d7ef01b3ac8dc79fd3e336041bc

Observation 80dd1ef7-0309-4364-90ea-9df1c482d955 · outbound

This paper cites Deepstack: Deeply stacking visual tokens is surprisingly simple and effective for lmms.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Deepstack: Deeply stacking visual tokens is surprisingly simple and effective for lmms

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.151044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.151044Z digest=sha256:9a632b4228a28c81f8b13a9b593976a18c248d3b7545bcb72ba6ac254cab0879

Observation 0d68cc44-7ddd-4136-8f03-60e9ada92ccc · outbound

This paper cites Reverie: Remote embodied visual referring expression in real indoor environments.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Reverie: Remote embodied visual referring expression in real indoor environments

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.159862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.159862Z digest=sha256:93d406a6c4beeb68ffc9b5326f6ed7f916d356fa7158ffe29820400436e2d14a

Observation dfa1ae66-9b8d-43fe-b109-05de296d5261 · outbound

This paper cites Learning transferable visual models from natural language supervision.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Learning transferable visual models from natural language supervision

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.170592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.170592Z digest=sha256:8abbd3e2b2b8af1a41efda65e37cb7fe826dfbe2a7344306a60588abe6e608f4

Observation 7d4bb9c3-83df-4640-be24-8a1c883918ee · outbound

This paper cites Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.177491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.177491Z digest=sha256:ce3065ddf0b937c81862ad1b524c45ef0f6beef2523cf2d6c7ed5b4c9acef959

Observation 8f7755a6-a814-45d4-b3b8-f6871d6fb957 · outbound

This paper cites Language- aligned waypoint (law) supervision for vision-and-language navigation in continuous envi- ronments.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Language- aligned waypoint (law) supervision for vision-and-language navigation in continuous envi- ronments

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.183795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.183795Z digest=sha256:83a89764dcdd65041959fa9998366deb2f5f9d380d47d83757eea2731b17a3d9

Observation 0d9ce42f-e30e-4f61-a6a1-c1de52d2c815 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation A reduction of imitation learning and structured prediction to no-regret online learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.190707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.190707Z digest=sha256:ce29c5ee9d04730d495db17e47071bde76c85c98ee3105b2a13806f4c13df3ab

Observation 13fd62c1-a987-4850-a83e-dbb1fbd4111b · outbound

This paper cites Habitat: A platform for embodied ai research.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Habitat: A platform for embodied ai research

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.199713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.199713Z digest=sha256:1781df30cfa86093bcbf90edab241f03f9267aabfa457723c9924b3b3f580689

Observation 84ea8eb2-66f5-4c3b-a52c-24283fcd70d0 · outbound

This paper cites Learning to navigate unseen environments: Back translation with environmental dropout.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Learning to navigate unseen environments: Back translation with environmental dropout

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.210468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.210468Z digest=sha256:21b421e58aa02d440d5a63eac86fca97662a64a0a547f04f2c13f2ea82ffe762

Observation 3770f99d-138f-40ca-bceb-704b747b0121 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Gemini: A Family of Highly Capable Multimodal Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.226291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.226291Z digest=sha256:5d501d490cded324df110151da25eb721edf29c5657d9ebd8e57a09398dcda62

Observation 10b9fe1e-98db-47c8-bd49-2c27f023c4c8 · outbound

This paper cites Vision-and-dialog navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Vision-and-dialog navigation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.238540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.238540Z digest=sha256:03992aae6fd66576f4436b7bc4b6c63e50764503ca37b6f292ac42f9d8a06af7

Observation 591f0d4e-cbb9-43e2-a3ed-f029809c1d62 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.245014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.245014Z digest=sha256:7e36031329691f96173953c6b0d51269ee2be75858908d9a8f7cb99337a703e4

Observation 08aaa791-0068-497c-8215-4577669021dc · outbound

This paper cites Dreamwalker: Mental planning for continuous vision-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Dreamwalker: Mental planning for continuous vision-language navigation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.256233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.256233Z digest=sha256:aa650852ee5b44bd933c017bbfa6409b20be8cc1c672bf2ec7a2017fe25da273

Observation b07c4916-dd03-4954-9ff0-9a49313f3033 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.265665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.265665Z digest=sha256:726390acbcddf2533be9ca17d18419a6e1dfc063979b995083b8496a4b530f8a

Observation b9ed50af-392c-4949-a1fb-6d13ddcc4585 · outbound

This paper cites Gridmm: Grid memory map for vision-and-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Gridmm: Grid memory map for vision-and-language navigation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.273325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.273325Z digest=sha256:75d2ab3e0dedc854e01bce19a238966a3ba5f32fade5d50d7b97c03bbb8e1b34

Observation 96636e53-13ff-47ef-b6f8-59243aed4443 · outbound

This paper cites Lookahead exploration with neural radiance representation for continuous vision-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Lookahead exploration with neural radiance representation for continuous vision-language navigation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.280485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.280485Z digest=sha256:672378e8d5295c9c67c3efbe18da49911a8f3198fe88ffeb063ed3617d7afc1c

Observation 0ff8670e-d957-4cdc-b122-94b3404acb98 · outbound

This paper cites Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.289618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.289618Z digest=sha256:15eecb269959cc81424279ca94823d6c57725faf0b3710f32dd37f7b6422ab4b

Observation 997c602b-1ff2-4595-b5b4-bfde57e93136 · outbound

This paper cites Scaling data generation in vision-and-language navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Scaling data generation in vision-and-language navigation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.300070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.300070Z digest=sha256:b9e4086ef4b0892875db23d6d5e2e69478c74b1e4b9ab37b7b56482ef7bff4cf

Observation 5bd1de4b-18ff-4341-9a0f-39851b05a53f · outbound

This paper cites StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.312697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.312697Z digest=sha256:8cf6cfb4f1f1dd432e9287047d1d6b334cf09c21d648232d781451c131454e32

Observation d85f88cc-a354-4a16-871c-53d9a16153ca · outbound

This paper cites Gibson env: Real-world perception for embodied agents.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Gibson env: Real-world perception for embodied agents

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.320554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.320554Z digest=sha256:e3c51db208bafb78307da7e25e02130ff4173af70bdd623100b38ccdfce5b3ec

Observation 62381cda-58df-4449-8cfe-7b594c6a817f · outbound

This paper cites Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.329176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.329176Z digest=sha256:54d7d06db2320cf250d1fe656aaf2123717adbf4f1ec3f2af9fadc4c72ec5ecf

Observation 9018df27-be93-47a9-a72f-0f7de330cdc5 · outbound

This paper cites NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.335446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.335446Z digest=sha256:168d8563f6e91eab97cdee5bdc4e9ef0dedf49aa0eaa5a251ace9d01c4d9f1a7

Observation c87b155e-1dd5-4476-8cec-141e67f181fa · outbound

This paper cites Lyra: An efficient and speech-centric framework for omni-cognition.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Lyra: An efficient and speech-centric framework for omni-cognition

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.343242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.343242Z digest=sha256:2d18baf3c038ebefe8897476d3f37410850598fcb211ea80418bda6a734b3d2d

Observation bd217528-721c-44f4-9ff0-27980426c0c5 · outbound

This paper cites Navgpt: Explicit reasoning in vision-and-language navigation with large language models.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation Navgpt: Explicit reasoning in vision-and-language navigation with large language models

Reference 48

Resolution
malformed identifier
no resolver link, observed 2026-07-30T20:42:04.352249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.352249Z digest=sha256:a63802b3cfdb2067a941a09505856ba430c9fb75a076f4fe05f22aac7fd41216

Pith citing papers

No inbound Pith citation observations are available.