Pith. sign in

Paper Citation Record · LEDGER

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

As of 7 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 54 inbound Pith citation observations for arXiv:2507.05240.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.05240 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:37:06.534556Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 54 of 54 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:31:40.601037Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T02:27:48.997495Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1a6ac6b-8bb6-48fb-89a0-709f99a47467 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.223688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:04.300000Z digest=sha256:aecc55fc18cb262f96a274e4f9dc0f22569bb291e7fae389ceb12b59c6191113

Observation 6541ff71-3000-4d95-8b2f-ac2d56703970 · outbound

This paper cites Zhang, K.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Zhang, K

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.210933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:04.451288Z digest=sha256:ce9f7936775aea59b9dcc6a335c5123ca0d45161a0b1f14c3284eb7f7f3fe3c7

Observation d8ab1ff1-3102-4214-9e23-0689b4146f99 · outbound

This paper cites Zhang, K.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Zhang, K

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.197826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:04.485778Z digest=sha256:7d5dde490bca562e140963a38aedc917bc95323019580407593659a2ce8d688e

Observation bd0a7da3-29dc-46f1-b4ae-dbdcdbb2ae70 · outbound

This paper cites Cheng, Y.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Cheng, Y

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.185208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:04.590753Z digest=sha256:1a8bd7c77b59198d2b3fa0211e0583597b72e8f52b4ce278bdace8f168792cd8

Observation 09fa0006-9fc2-46eb-a32f-6ff7111d3bf4 · outbound

This paper cites MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:04.732234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:04.732234Z digest=sha256:b428aa42a0154108aadc47d9583cedb1094f5b61d4d05e3a9cca7e5945a168ad

Observation 72e2793a-74b2-4fd8-a31f-1ae78d54c75b · outbound

This paper cites Anderson, Q.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Anderson, Q

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.172347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:04.950494Z digest=sha256:b6a5e6e033ff317befa883281ce2cb124c4c1368649af77884cb39974ac82e58

Observation c253d784-c227-4c25-b950-f10c9c8ce7cd · outbound

This paper cites Room-Across-Room: Multilingual Vision-and-Language Navigation with Dense Spatiotemporal Grounding.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Room-Across-Room: Multilingual Vision-and-Language Navigation with Dense Spatiotemporal Grounding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:05.110483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:05.110483Z digest=sha256:b7bdfacfdb37ef2327434d83da33411b7e15bff6d01b1881a0bc908b8ce927c7

Observation e55fc85a-52af-4434-9ccb-1b5815afd79a · outbound

This paper cites Chen, P.-L.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Chen, P.-L

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.159661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:05.226968Z digest=sha256:bc4a8c216b6f5a47a80f171958b27e64624fafe5e4d0a43a568625e65badee7a

Observation 47ced714-6e49-4960-8ae1-df75802caca0 · outbound

This paper cites Chen, P.-L.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Chen, P.-L

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.146748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:05.357839Z digest=sha256:443465f00eafb1a76d7e806147f867b7be1337d5346f7c5f007e771f00afa7ff

Observation bd239d17-36cb-499b-988f-b31536916bd2 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.133373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:05.600830Z digest=sha256:dfe56e205b41aaf76685ca60d9487365633c9bd88db716c6a88ca07e77bbe90b

Observation c91a7d8e-19bf-4711-9c03-293b3c2b4026 · outbound

This paper cites NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:05.744748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:05.744748Z digest=sha256:a96390cdbd65e87c8ca0ea48097b4c3b7e1eb6f0cc27bd951bc2282726431dc1

Observation 8baa8afb-764b-49e8-a832-da8750fd07b4 · outbound

This paper cites Language-Aligned Waypoint (LAW) Supervision for Vision-and-Language Navigation in Continuous Environments.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Language-Aligned Waypoint (LAW) Supervision for Vision-and-Language Navigation in Continuous Environments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:05.891661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:05.891661Z digest=sha256:39603e99b055744d8481984b05f0e43c8e12308821b94a635c89e029033f7651

Observation c78712a0-9dfe-4b68-b2e1-6b2f29e90d6c · outbound

This paper cites Georgakis, K.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Georgakis, K

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.119892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:05.982383Z digest=sha256:0679eddfff7271c7c38dcaea6174774f9974a18b79bbc256107a433b213cf996

Observation c5a8b931-cc68-43f9-8a96-7026892daff9 · outbound

This paper cites Weakly-Supervised Multi-Granularity Map Learning for Vision-and-Language Navigation.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Weakly-Supervised Multi-Granularity Map Learning for Vision-and-Language Navigation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.054748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.054748Z digest=sha256:34605fa4b43d9b95f231c2b6e0fe206c6ae69fff309dc59b25f8655c6829d941

Observation 25ee39d9-2525-4f94-89b5-442d9682cb04 · outbound

This paper cites ETPNav: Evolving Topological Planning for Vision-Language Navigation in Continuous Environments.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling ETPNav: Evolving Topological Planning for Vision-Language Navigation in Continuous Environments

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.202497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.202497Z digest=sha256:db0c758e6f14fbdeff65b3048cd49cbb837525e40fae20ddd0d97cc4fa1aed64

Observation 59db7e1b-557f-41af-a63f-7e4b1a083fa3 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.107134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.360697Z digest=sha256:72dd16563e4d28a51a806d3b5c7c6b211cead6449415fe39db06b151caf647c3

Observation 5fd12fe7-dcab-45d2-b277-06025467a4f8 · outbound

This paper cites Krantz, E.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Krantz, E

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.093534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.399558Z digest=sha256:30db9bf34e6e986b34d3e2b891e63b7ddb1c1e66bd6702b715e084ece08bf113

Observation a4fca432-dfc3-49f0-9d7c-0d54cb817fe3 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.081039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.404949Z digest=sha256:1ed4f760ee84994ba6aca9496c2f69c06ac13ce1a1c7d1c0c765bd4bf4c30ba7

Observation fb5b9c73-7622-40dc-9f13-d8c9e064af3e · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.067772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.410064Z digest=sha256:606e3552bede8708e4313c4f7d140d358a1953a9b6748f025a7db07c3dceeccd

Observation 4f84e770-f257-4bd1-a943-4b38c6ce019d · outbound

This paper cites Discuss Before Moving: Visual Language Navigation via Multi-expert Discussions.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Discuss Before Moving: Visual Language Navigation via Multi-expert Discussions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.415188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.415188Z digest=sha256:e41b6e7a768827ab060436a0a63e7738b947b71b1296041001386e3a89b7f872

Observation 7320402f-b533-4006-9ce7-6023d1c394ab · outbound

This paper cites $A^2$Nav: Action-Aware Zero-Shot Robot Navigation by Exploiting Vision-and-Language Ability of Foundation Models.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling $A^2$Nav: Action-Aware Zero-Shot Robot Navigation by Exploiting Vision-and-Language Ability of Foundation Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.419946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.419946Z digest=sha256:aac705766d33584610adb87150f54b9803e485789a5a4b0af46044754d0df695

Observation b8e208ea-9c08-4139-bbb5-6828d2db7628 · outbound

This paper cites InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling InstructNav: Zero-shot System for Generic Instruction Navigation in Unexplored Environment

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.424089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.424089Z digest=sha256:e05db2675aa850519b2b458d787a39278a44b98927491107f24cccd16ab03af6

Observation a4aba587-729e-4d5c-bdcb-2de995417c69 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.054767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.428506Z digest=sha256:270e9c7ef8b835e1b8bfdb6121e7ce0560a0ced19904f739148dbb0e58a0aaf4

Observation 5a93a66b-de4a-4294-a7cd-492059cc3678 · outbound

This paper cites Chang, A.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Chang, A

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.042373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.432342Z digest=sha256:c1f7a4b8e8509bd95012857d78051e7c05513c879d560b01cf7f2480029b7883

Observation 4c8b3b58-0b5b-4be6-8eab-b55a3ff26e0e · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.029428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.436696Z digest=sha256:ebc1eff34c925103e0c436c176d82c279f81884ce14a8f535ccf04f875996e2c

Observation 69579633-7b64-4fe9-85cb-b9a9a3bd7d20 · outbound

This paper cites Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.440919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.440919Z digest=sha256:2d51a958a674c0f4a9838a39a6d637f2c666c02946bc796ee80a9ed0d6ec3e2c

Observation 8397d883-9741-4e98-b261-c1d90d60b90f · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:07.016919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.445428Z digest=sha256:10a1d25120e0ad6e9f215e74bd630a94e7d8e93ece5440d4fe7bb886acc3df39

Observation f86124a1-b5ac-4052-a9e4-c56c69400990 · outbound

This paper cites LLaVA-Video: Video Instruction Tuning With Synthetic Data.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling LLaVA-Video: Video Instruction Tuning With Synthetic Data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.449859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.449859Z digest=sha256:66cf5cf4bba001ff20039732ae60f06dea479f0e62e238d99bc8cf806c5e7550

Observation ebf53f03-e662-4614-8813-84e80b6c4ada · outbound

This paper cites Azuma, T.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Azuma, T

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:07.003856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.453764Z digest=sha256:25a4b3c197dbb61d9e774a89c79fb4c83b03517c73053cd0c05fed7407cf725e

Observation f7ca8001-b5f3-4738-bb41-372f06b969e9 · outbound

This paper cites Multimodal C4: An Open, Billion-scale Corpus of Images Interleaved with Text.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Multimodal C4: An Open, Billion-scale Corpus of Images Interleaved with Text

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.458067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.458067Z digest=sha256:0c457d2f723f297b54bb24a6722cad78009dec6f5bd68f54bf698a5963ff233b

Observation eb6eac91-2bc6-4252-84fd-0a0a7a3ccc65 · outbound

This paper cites Qwen2.5 Technical Report.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Qwen2.5 Technical Report

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.462400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.462400Z digest=sha256:38480922695ce10b34cb536a77c74b3575f08d4f228b265ca8d57e964abb775d

Observation 18acfa32-a0b2-432c-9294-f0599be06795 · outbound

This paper cites Krantz, A.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Krantz, A

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:06.991080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.466678Z digest=sha256:93ead1fe87d2bb7d1d69968ee9366558d4ee3fa21caed8e06e3791fca4a0561f

Observation a0e04274-257d-4e3f-b352-7784d23846cf · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.978424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.470509Z digest=sha256:e77f1261925054bd19ee95e985b557b6d6bdec801285f895750085f20d5c6f6d

Observation e8e866ee-2d88-49dc-b40e-8dbad80f2a30 · outbound

This paper cites Krantz and S.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Krantz and S

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:06.964538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.474563Z digest=sha256:a26d198a0878c10445b3703a55e2812979ce5d179d55957cffc05e36bbba91a1

Observation 76387e02-7a57-4976-8f3a-6bb42d3bbf30 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.951187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.478494Z digest=sha256:ca5d60ce7fe0f91c2501bf265a048b88a1255991ff2885ab7b1516e08c3b3356

Observation 5dc7463b-a2d4-48bb-91eb-e05ab46f3b27 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.937696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.482477Z digest=sha256:211299c2c21a9469bc25e6340be03c43105d5497d886a4409abb7ed760dc9a74

Observation bde7ccd3-c346-4362-900d-b5355b079ea7 · outbound

This paper cites Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.486394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.486394Z digest=sha256:1a2e05babf8d23708df590ac51a5e3cb89a48764e50d805f2bc577ec53426219

Observation a883a61c-328e-4ccd-a7d9-b87b7e9eee6d · outbound

This paper cites Krantz, E.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Krantz, E

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:06.923597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.490394Z digest=sha256:c016dd9b8b4b87791fb5c15dadb46aa59d49d731f8a861877b7c21ecd0d3a69a

Observation a8ff1c39-a971-4b8d-8343-8f85472aeb0f · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.910528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.494300Z digest=sha256:3a39271d78f0ff283bed2d6d3e6a44eaf6a9b51187d1882f27b3697c7d320163

Observation 63d8b9eb-6a4d-4eda-96a0-45a75d1b920e · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.897004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.498008Z digest=sha256:7a2bb8160eb458c29e2815c20d3fa90f6f688b4c07f2325fefc4826172d8af57

Observation 8b2c080d-c8e8-409e-b21f-55da58bc9b73 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.882131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.501769Z digest=sha256:6a3410f5117dec5a7c18a935e5121fb0cb3f30a153f5aebd4e8519eb9a0a9618

Observation 5161b1a4-7d3b-4786-bfc1-00d966709b5a · outbound

This paper cites An Embodied Generalist Agent in 3D World.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling An Embodied Generalist Agent in 3D World

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.505642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.505642Z digest=sha256:6059a4f7ab634509b28b8ae8e73c3359e8835051402588b5081b3e0ce72409ba

Observation 88a6554e-19af-47e0-8da7-d9bf491cdb57 · outbound

This paper cites Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.509669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.509669Z digest=sha256:f48dfa4319fb7eae85143e51625f739a06186ae0ea442ae5a9a55cdd3c50583f

Observation 8cb2eccc-24b5-42f7-a119-9f2e977d26ab · outbound

This paper cites Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.513506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.513506Z digest=sha256:df37d0947c595e93bb035052af147628913a5a8dbb9e8d3f5d68d267d04a3b22

Observation 840b945b-14ad-4537-8da6-ce1e8e4e14f0 · outbound

This paper cites Zheng, S.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Zheng, S

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:06.868293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.518184Z digest=sha256:93366121a3dff3fbc89640d05a635f0369672cc8bfb2717e3ec0915eb48abaf2

Observation a8c54734-3622-487b-a795-2e0dfe7eaba8 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.853571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.522164Z digest=sha256:1665bd19939448a4ea454ba2533738f53d915ae1a606d40c5ac4add64fd8e3c9

Observation 265d71aa-7a0b-4165-86b6-a9ed00351bdc · outbound

This paper cites ESC: Exploration with Soft Commonsense Constraints for Zero-shot Object Navigation.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling ESC: Exploration with Soft Commonsense Constraints for Zero-shot Object Navigation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T19:37:06.526119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:37:06.526119Z digest=sha256:6eea271b3aabacc52e1cc440d684e54b55c8412d14df576fa9f2804dbed5e35b

Observation 201e26d8-a14d-4ac7-a3df-150ddb500020 · outbound

This paper cites an unresolved cited work.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:37:06.840670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.530512Z digest=sha256:0fca227e6b069077c47ab7120fe7bf57fa3211ab150534306c190a31f91e48bb

Observation 34c73bbf-3223-4720-a970-82b31b304232 · outbound

This paper cites move forward 25 cm,.

StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling move forward 25 cm,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:37:06.827433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:37:06.534556Z digest=sha256:838328c91fb28a181a17b2a76d4a41f643a18ec44f4159313745df0bc89ccc12

Pith citing papers

Observation b2de16a7-714b-4211-9327-41f774905747 · inbound

CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model cites this paper.

CorrectNav: Self-Correction Flywheel Empowers Vision-Language-Action Navigation Model StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:09.130917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:31:09.130917Z digest=sha256:ac457cd7a4e18b1c2de143e75cdca0b17b18a641b4ffc4eaef67102cf1670d8b

Observation 4587285b-ab9f-4591-ad7c-fed4f4c6fc10 · inbound

Nav-R1: Reasoning and Navigation in Embodied Scenes cites this paper.

Nav-R1: Reasoning and Navigation in Embodied Scenes StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T17:31:43.003448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:31:43.003448Z digest=sha256:dde38c8a9ca13adaa67b22944d7b871076ad0fe90e4f5497e818cc94d8271e16

Observation 2c49428f-c9e1-4d24-a360-30e4c71bfa2a · inbound

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation cites this paper.

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T21:02:38.013115Z digest=sha256:bd397bd24c2eb0c253c62ac7df72e6bbaeaabe3317f05a942da179995760217b

Observation bf0af0d9-9606-4359-817d-51f9ee3a1dee · inbound

Efficient-VLN: A Simple yet Strong Baseline for Efficient Vision-Language Navigation cites this paper.

Efficient-VLN: A Simple yet Strong Baseline for Efficient Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T17:18:19.601385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:18:19.601385Z digest=sha256:b940d4177e55d58d4cdb2e9b7a95682b1f213c02d2d0bbf905d8fcd94a06fa61

Observation 3c942637-7af2-4729-8768-4c1b6cbe1b21 · inbound

AstraNav-World: World Model for Foresight Control and Consistency cites this paper.

AstraNav-World: World Model for Foresight Control and Consistency StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T19:23:58.769472Z digest=sha256:c115ed8ef289232335517536c81be541ee6c5b7f15d83cc17e8106c16b7b21cf

Observation 8d21f7f6-9181-49c8-8a7e-b89e56fcd891 · inbound

OpenFrontier: General Navigation with Visual-Language Grounded Frontiers cites this paper.

OpenFrontier: General Navigation with Visual-Language Grounded Frontiers StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T11:36:24.334860Z digest=sha256:bf65a3e651d797700e58901771672bd4bb33daada5ba38a5aa63be8a52b406bf

Observation 14dcf89f-d174-4c3c-83f4-3cbe16fefee9 · inbound

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness cites this paper.

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T14:50:58.269817Z digest=sha256:f7b05a01468603cb199f544ac9ebd41291e73ba728d108a405ca4e0597ac32b6

Observation 217214d1-4666-40e3-87d4-c293490d4e9c · inbound

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs cites this paper.

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T10:41:47.297072Z digest=sha256:faae4c6c835fb1a6b44553f194bb0343a5e27c6c2ebfc019635bc1310920fdac

Observation 653e5bef-c4bf-4433-9033-1965d98b3020 · inbound

Structured Observation Language for Efficient and Generalizable Vision-Language Navigation cites this paper.

Structured Observation Language for Efficient and Generalizable Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T17:15:42.466280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:15:42.466280Z digest=sha256:438831ad4e8b2726952eeae8da1d5a0e3afdfc094c8de76eeb0e30e2dbd9d97f

Observation eb6c8ea4-5e11-47c8-acfc-c1a46dae0cdc · inbound

Structured Observation Language for Efficient and Generalizable Vision-Language Navigation cites this paper.

Structured Observation Language for Efficient and Generalizable Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T05:43:12.553804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:43:12.553804Z digest=sha256:5c66dee69c2448621c805b3e421e88e3cf05fd6bfc04909fc6aa545c00d3f49b

Observation 6ef8e15d-04bf-4dc9-93fa-63b2bf046875 · inbound

FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation cites this paper.

FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T08:13:12.029402Z digest=sha256:a9c2a2831236e8b373317f25f4cbe6513882118721a01732b5e746593219abd4

Observation 45cc06b1-8688-428c-83f8-b1b4ec15f1c1 · inbound

Think before Go: Hierarchical Reasoning for Image-goal Navigation cites this paper.

Think before Go: Hierarchical Reasoning for Image-goal Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T05:43:27.972164Z digest=sha256:c1f76dc7973f81b9e952e9496929fa6fcf41fcea88c095bc06bff4eef561492a

Observation 3c48c546-de92-4b44-ac51-701dad9ccf18 · inbound

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation cites this paper.

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T06:50:34.310831Z digest=sha256:320451543468de173be90084f90f8bd317ce4a19dc9131396e1f272d3eed2ea3

Observation faaa55f3-d870-4c97-8899-05c840f53efe · inbound

Existence of small semi-vortex solutions for the cubic nonlinear Schr\"{o}dinger system with Rashba type Spin-Orbit coupling on $\mathbb{R}^2$ cites this paper.

Existence of small semi-vortex solutions for the cubic nonlinear Schr\"{o}dinger system with Rashba type Spin-Orbit coupling on $\mathbb{R}^2$ StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T18:51:26.143516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:51:26.143516Z digest=sha256:1b27e8b01d71e11bdd094ad86f6d3f169e531fcb54534c95949799c8c23a9846

Observation ab54c9d6-c0e0-4d7c-a30c-e1d71dd5c6bc · inbound

LiveVLN: Breaking the Stop-and-Go Loop in Vision-Language Navigation cites this paper.

LiveVLN: Breaking the Stop-and-Go Loop in Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T01:51:23.379734Z digest=sha256:b24d269ae6b205ba4ab588c910746575d2fd4dd4b127a12c91c2ca3cde658ffd

Observation ade1d0c2-2fbf-4c02-afb6-6cd0b1c7e553 · inbound

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching cites this paper.

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T03:10:29.477937Z digest=sha256:a685aabf18ac6206d6ba536b8b0acde1f48d664e70c3fea1bd96fd8a6fa41676

Observation b7da6bd7-da02-4781-96ea-010cb130f03c · inbound

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation cites this paper.

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T05:05:33.606975Z digest=sha256:8e0a6a7fdd8388b77268db993f7f7f3b3e034044b0cafc9f6d75c5299553ea5a

Observation 150b453d-f45f-4709-a1c2-f906d238b095 · inbound

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation cites this paper.

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:04:57.627160Z digest=sha256:3444c551912eda285ffc92f47411f760152960969fc858a7137472fa8f5e841d

Observation e2fdc933-7aa6-46ed-acd9-72cb88c729ee · inbound

Beyond Isolation: A Unified Benchmark for General-Purpose Navigation cites this paper.

Beyond Isolation: A Unified Benchmark for General-Purpose Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:33:53.557357Z digest=sha256:7dfcf46158cd19be2cbd96577a1168b36876f7cdf789b40a4c7c23f8a5f61fdb

Observation 191a073c-a88f-4214-9dda-cbe1adb50dc3 · inbound

Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation cites this paper.

Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T05:25:18.120832Z digest=sha256:8d1961d401a673b1f17fa0e32aa600a1940f34b8630222a7243d646ae21cbce9

Observation 741707e9-0f07-432f-a0f5-fa3b721e0f8b · inbound

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving cites this paper.

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:54:50.887381Z digest=sha256:6e584840d1b7d94698aac70f5b0a56e7eb355fed1e4c6e9c1d358aaef35992b6

Observation cf1bd163-89d5-49d5-9956-379b7a459278 · inbound

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving cites this paper.

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T05:07:58.953866Z digest=sha256:6757ddfb34339aa0bac481bb423988891208db12b51bad93de1d512288db5572

Observation 45a9a7de-3aa9-4807-82cb-4ab099d64764 · inbound

PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World cites this paper.

PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:40:59.877854Z digest=sha256:9b02bdf7d45ae769898ebdf023336d0186b3904d96aa42b30725dae907e3f294

Observation 42e39c8d-eacb-47f6-ab41-c6681a665456 · inbound

PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World cites this paper.

PanoWorld: Towards Spatial Supersensing in 360$^\circ$ Panorama World StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T16:57:03.172340Z digest=sha256:ae62dcd5c7573205ff65c67eec8ea235ad9ddaeed8018767c0c5b94b5a0aa09d

Observation 8fa5b160-d3d0-4b95-9686-6349ee77f524 · inbound

What Limits Vision-and-Language Navigation ? cites this paper.

What Limits Vision-and-Language Navigation ? StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T17:59:58.891151Z digest=sha256:c42e85ea4b720cf0f7a275030f37faf34b20bedd5ed3901196016df55daf1668

Observation c37c7315-a25e-484d-88d1-25b4733c21cb · inbound

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation cites this paper.

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T13:31:16.012419Z digest=sha256:a2725e13f78a73f298d15e588e42f6614bd47bd84d07ed9e5875b4467eb90ee6

Observation a6e91a8e-0f90-4f9d-8334-fde2c0916773 · inbound

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation cites this paper.

SEDualVLN: A Spatially-Enhanced Dual-System for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T19:42:41.238072Z digest=sha256:17cd5708af87a05e17bbdcb0f85d15aec57beaf77f2ad4d12c17f5b6336474be

Observation bd15211b-f171-4fcf-840b-4a377359cf42 · inbound

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation cites this paper.

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T06:45:35.920991Z digest=sha256:1eceb27986375c6600d261eff756c67430412dd33dc9df305fda076c61906507

Observation 134d7494-b0f0-4396-bf3e-498cb0861b5b · inbound

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation cites this paper.

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T04:46:03.020800Z digest=sha256:cfa15b33c26bb898149bde2b232fd866f315c7d0ab1ff41cbe27c30472707c68

Observation aa6fd94a-5697-4095-93cc-0feb4ef7d4c4 · inbound

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation cites this paper.

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T17:01:20.073666Z digest=sha256:80b42373d23de0742e4a33b7fce4294cc73ee5dc35591993d3a285ec059bd1b1

Observation 09c6ba14-c83e-48a0-9560-839fbac711df · inbound

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation cites this paper.

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T11:48:14.888295Z digest=sha256:18571dfe8b20e944733356aeb05ce94b7eb0ef56177cb74bececa5b4444e7022

Observation b1d706e1-3f63-4d29-80d9-56198028d780 · inbound

TARIC: Memory-Augmented Traversability-Aware Outdoor VLN under Interrupted Semantic Cues cites this paper.

TARIC: Memory-Augmented Traversability-Aware Outdoor VLN under Interrupted Semantic Cues StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:09:07.785292Z digest=sha256:cd50820be6a66ca737f6c96e48e81813798cd4783a21e7bb57d81e4e20e22fc2

Observation 71d33fdf-bd20-4dd1-81b9-be4ff0fcb5aa · inbound

OneVLA: A Unified Framework for Embodied Tasks cites this paper.

OneVLA: A Unified Framework for Embodied Tasks StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T17:05:08.124096Z digest=sha256:f44a5e4b42958142b503ec48e9a4c4cdc42bf5379073c223f2cdc7939be81c08

Observation 116fb1b3-902f-48a1-9c56-778e009d468b · inbound

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation cites this paper.

Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T15:39:13.803571Z digest=sha256:48899acaa38ae8f426a2a7f720657fb1791b1fc1258d8ca3298c223d3313f7f6

Observation 5b167f9b-3fd9-4446-8918-6df6fd5abdab · inbound

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation cites this paper.

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T21:40:58.329546Z digest=sha256:b35560e41b8c21bf4a1df2890fb9aa9b41f127fbe672b80c176560e691dca251

Observation a4170dc7-cf4a-4b14-ac9b-a2763b876e97 · inbound

Slow Brain, Fast Planner: Latency-Resilient VLM-Augmented Urban Navigation cites this paper.

Slow Brain, Fast Planner: Latency-Resilient VLM-Augmented Urban Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:16:52.173943Z digest=sha256:30b6bda2ebb8657a80646d687e8caeb4e0a9b0b244c6da23cd5fff11c06b9c5f

Observation a0b2007b-4438-46a3-9b5d-3a65a008b512 · inbound

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks cites this paper.

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:42:23.040915Z digest=sha256:0e15b8f90de3125166d04c661d07177fb105587ae7b0e3343436534526b770f6

Observation 194dc1f9-4351-4676-a226-02076be9dec7 · inbound

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation cites this paper.

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T05:05:56.065261Z digest=sha256:979db8e43fefcada988acc1e7b3e3693c1698ae8e25ba679d4a0bf496b656523

Observation 39bd6d97-f81c-46e9-a85a-6b0c9337e069 · inbound

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation cites this paper.

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T14:04:46.400336Z digest=sha256:c55079fa323f793bf9b9985a4d37058ade022c8c88bd88b328de0fa2f0ebf9a3

Observation 3c0c299d-7885-46f2-a820-2c75e7b79f35 · inbound

LIME: Learning Intent-aware Camera Motion from Egocentric Video cites this paper.

LIME: Learning Intent-aware Camera Motion from Egocentric Video StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-10T01:18:45.609520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T11:01:22.425665Z digest=sha256:db61db8764e8c3ade29f89546bffa0c5075d99689ac918620fb90a1582265199

Observation eb84062f-aa1c-4a85-998a-0afa452ba9d4 · inbound

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI cites this paper.

ACE-Brain-0.5: A Unified Embodied Foundational Model for Physical Agentic AI StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-11T19:16:57.396710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:16:57.396710Z digest=sha256:fc30d18175c636ba492c2e945b4d272eaac83f5d34963a834dd9ae54a13e0448

Observation a1a82472-770e-4e09-bc0a-609bd5ab1617 · inbound

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator cites this paper.

Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 92

Resolution
verified exact
local_arxiv, observed 2026-07-11T02:27:49.014528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-11T02:17:51.026130Z digest=sha256:da6ecd9b085e6b2cc950a16712a25c01483016dd337984951927ac5d8be838cf

Observation be41978a-e6a9-469a-a9c0-b2f8c2f53114 · inbound

A Comprehensive Survey and Systematic Real-World Evaluation of Embodied Vision-and-Language Navigation cites this paper.

A Comprehensive Survey and Systematic Real-World Evaluation of Embodied Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T15:39:59.878169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:39:59.878169Z digest=sha256:d4722aaa2dfdd2337c00dcc97c790748d853360980c6580e810c0e18990d5560

Observation 5d82446a-78e4-4511-94a6-5cd256284c50 · inbound

ABot-N1: Toward a General Visual Language Navigation Foundation Model cites this paper.

ABot-N1: Toward a General Visual Language Navigation Foundation Model StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-14T12:10:21.115628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:10:21.115628Z digest=sha256:b5fe83abf8f2e152474527c1e5bb33b79dbeefca3256cf6b69cc53e4b8cfeca6

Observation 35ace54b-82f0-409c-bcfd-8e99ef9ab06b · inbound

ABot-N1: Toward a General Visual Language Navigation Foundation Model cites this paper.

ABot-N1: Toward a General Visual Language Navigation Foundation Model StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T07:19:43.503753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:19:43.503753Z digest=sha256:356c786056f975b4dc12344e1841fdda97309320c584c18ec9d9dba23ee6bd2d

Observation 9d75388c-2b0f-4b72-9303-ff1331f411e1 · inbound

Joint On-and-Off Policy Learning for Vision-and-Language Navigation cites this paper.

Joint On-and-Off Policy Learning for Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T05:13:53.937100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:13:53.937100Z digest=sha256:60a6b3eeb9982a81d27fc6ff61eeb07897fa1cf8f9bb9a4d6aa87f08e044c23f

Observation e02f8ffc-b5f8-44a5-9f3e-ad1db1649f85 · inbound

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation cites this paper.

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T19:56:51.921232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:56:51.921232Z digest=sha256:c09809dfcd448371f13cac1e99369a73be80787ac38f0646b1f85fc1376055b4

Observation 5d51fa7c-e599-4faf-b223-5cb276df9a49 · inbound

Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation cites this paper.

Anticipate Before Acting: Future-State-Conditioned Vision-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T16:20:01.321694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:20:01.321694Z digest=sha256:5bbcf40c8ea1451b151cb9b507191d8b7ac12df87e414e8251be78f75351fb9c

Observation 7a4571ea-6daf-46cb-a412-8ac04f682ce0 · inbound

NavVerse: Benchmarking Indoor-to-Outdoor Embodied Navigation in Continuous Robot Simulation cites this paper.

NavVerse: Benchmarking Indoor-to-Outdoor Embodied Navigation in Continuous Robot Simulation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T12:04:39.750456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:04:39.750456Z digest=sha256:ea16b52a7642ae98f07b1bcb011b28d1e57ac5640558703559dd04e165ae7b3c

Observation 155c7167-7f48-4ec5-9a83-bbcf0dfb6a7c · inbound

ReferTrack: Referring Then Tracking for Embodied Visual Tracking cites this paper.

ReferTrack: Referring Then Tracking for Embodied Visual Tracking StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T11:00:24.818615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:00:24.818615Z digest=sha256:f4e0a98bce874adbd4562ee7f93ebdca069dc0540fbfd382bb925843684e59f4

Observation 5bd1de4b-18ff-4341-9a0f-39851b05a53f · inbound

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation cites this paper.

MemVLN: Episodic and Procedural Memory for Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-30T20:42:04.312697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:42:04.312697Z digest=sha256:89faf8c89ef909e7a44672d6653c1e160c2201b48accaf1de6703e97390baffa

Observation 0241c4c3-f28b-4c6f-adfc-235710884803 · inbound

Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation cites this paper.

Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T00:43:26.897230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:43:26.897230Z digest=sha256:2bc1f49975c7c757f024829b6335d316f69aff09a81ecbe03f3cb7e804d665a9

Observation 5a68a9fb-37e5-4bcc-a0fd-ab2db8827977 · inbound

Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation cites this paper.

Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T03:24:24.621076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T03:24:24.621076Z digest=sha256:707079c799eed2b530ccb2ab940eb421a588843eefd40bd8cd21c116fe38a609

Observation e8cc92f2-e08c-48c4-a469-0c38834bf70e · inbound

Think in Sets for Streaming Video Token Compression cites this paper.

Think in Sets for Streaming Video Token Compression StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:40.601037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:40.601037Z digest=sha256:6eaa08df8c552c0e7f1bf54611652916407834612607714675ff387f9782d492