Pith. sign in

Paper Citation Record · LEDGER

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation

As of 5 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 0 inbound Pith citation observations for arXiv:2607.23181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.23181 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T03:26:19.198670Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved68
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1543a9b6-d811-4830-8b25-842e710e9c9b · outbound

This paper cites Aligning cyber space with physical world: A comprehensive survey on embodied ai,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Aligning cyber space with physical world: A comprehensive survey on embodied ai,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.181242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.181242Z digest=sha256:b3d366338977d156852a099f35dcbcb1c9d6abfab21df55ca1b749ca3dce4537

Observation dad256b8-f2ec-4a80-8955-165b6e4720e1 · outbound

This paper cites A Survey on Robotics with Foundation Models: toward Embodied AI.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation A Survey on Robotics with Foundation Models: toward Embodied AI

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.242060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.242060Z digest=sha256:b37bab8f674ba003c23a42956a5e5e7cedf755c078ced188dc3edff842ba67be

Observation e0e1dfde-c087-432b-8f6d-f7777281eea5 · outbound

This paper cites Beyond the nav-graph: Vision and language navigation in continuous environments,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Beyond the nav-graph: Vision and language navigation in continuous environments,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.356671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.356671Z digest=sha256:af5bacfd648b9b5f06d08402d57d0ab5f542ac3b35609cd1497a68b4b66c23b4

Observation a6a320fc-374d-41f0-b2ef-1cd372d368dd · outbound

This paper cites Bridging the gap between learning in discrete and continuous en- vironments for vision-and-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Bridging the gap between learning in discrete and continuous en- vironments for vision-and-language navigation,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.435438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.435438Z digest=sha256:3984d9b71ea850cef4370d40bb717bc0502dd6d99cbe7fad22d2511e3b7e0bd6

Observation 9bff19d4-2fe1-418d-97eb-86d9b67621df · outbound

This paper cites Matterport3D: Learning from RGB-D Data in Indoor Environments.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Matterport3D: Learning from RGB-D Data in Indoor Environments

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.585421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.585421Z digest=sha256:ebe360f71b4f35c5abb7c256f9ae8969e90c1cb6054f3501404d9fbabc91f656

Observation d3a2d63d-db1c-490c-9ed0-1c659ffe8fb6 · outbound

This paper cites Towards long-horizon vision-language navigation: Platform, benchmark and method,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Towards long-horizon vision-language navigation: Platform, benchmark and method,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.714018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.714018Z digest=sha256:4deb42f768683c5d327c52b1c40f83ad9eee8cc5043bef92f15bb70043064141

Observation f92ee1f3-9552-4f6f-a86b-babaeffd1511 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.823905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.823905Z digest=sha256:c5e23c5be065423e2c5b145bc469a6045f2b47a818a597abc72d3589dcf4abcc

Observation ae2e058f-32bc-4639-b2bc-ab9a4561dd83 · outbound

This paper cites Perceiver IO: A General Architecture for Structured Inputs & Outputs.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Perceiver IO: A General Architecture for Structured Inputs & Outputs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:12.987863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:12.987863Z digest=sha256:56a1d7d59a2703a9b8054e7a36fd12de55c53168e0f39ffe2e1023c0f2aab1a0

Observation cfa80771-e8bb-47b1-979e-308e94bc8a20 · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Flamingo: a visual language model for few-shot learning,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.099845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.099845Z digest=sha256:e2af8c59e4b9264d01d063febb9d3fd91d6ffcdb5f870259d4021cf3718d515b

Observation 92eb84c2-2bcd-48d2-b0b0-7014200d7bd7 · outbound

This paper cites Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.213780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.213780Z digest=sha256:47461b097a83beb3a2b1216b6fa25b5f2b72e6d808d3253518c03b4152094257

Observation ecce2c54-2a46-4517-b5dc-6ba247e2a2ea · outbound

This paper cites Span-nav: Generalized spatial awareness for versatile vision-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Span-nav: Generalized spatial awareness for versatile vision-language navigation,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.354966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.354966Z digest=sha256:2d96df7a2536e69ef331b02816b19ca9c8eeda8ae046140ca500611ca9ef88ae

Observation c4480d23-03c6-4b56-8ccf-e2fd385735f8 · outbound

This paper cites Structured Observation Language for Efficient and Generalizable Vision-Language Navigation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Structured Observation Language for Efficient and Generalizable Vision-Language Navigation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.492892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.492892Z digest=sha256:b14b5823537ea90ee95f367eaf538804293bf82bcae7ac683b88969fc706a2a4

Observation 033cad07-65c0-4a57-b4ab-173e856692cf · outbound

This paper cites Navforesee: A unified vision-language world model for hierarchical planning and dual-horizon navigation prediction,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Navforesee: A unified vision-language world model for hierarchical planning and dual-horizon navigation prediction,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.640270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.640270Z digest=sha256:799970861c12a0874b125ab0d12bdf762d08fb060a6fab5a82aeb0bcd020b5de

Observation 36f748de-f012-4c69-85ab-0ec337fbda4e · outbound

This paper cites Mapdream: Task-driven map learning for vision-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Mapdream: Task-driven map learning for vision-language navigation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.838016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.838016Z digest=sha256:934c8db70694c955cecb2f97ff43f540d145bdc701e7a959d316830475b9f107

Observation a77e19a2-6188-49d3-bf39-d92c66494f74 · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:13.968141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:13.968141Z digest=sha256:492c486905f04abe06f7d84e9192ca811201384e6a5c0ddd4610dad76948c340

Observation 12ba6ac8-4c30-4505-a73f-5a78cd6c547d · outbound

This paper cites Room-across-room: Multilingual vision-and- language navigation with dense spatiotemporal grounding,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Room-across-room: Multilingual vision-and- language navigation with dense spatiotemporal grounding,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:14.134079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:14.134079Z digest=sha256:0c3afe5c5bd7a4982e522dd9a7524f0e4f5888ff457d956de2f817030c16d1ca

Observation af713ca6-51f8-4a43-bebd-0b40b56e6cc3 · outbound

This paper cites Cosmo: Combination of selective memorization for low-cost vision-and-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Cosmo: Combination of selective memorization for low-cost vision-and-language navigation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:14.311905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:14.311905Z digest=sha256:d5aa3b548bb9213d4e820b92aeb848b2802fcd548eca041baa4bfdf92efcdd57

Observation d71e7220-74a9-4f08-b44e-16a5d8089a72 · outbound

This paper cites 3d gaussian map with open-set semantic grouping for vision-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation 3d gaussian map with open-set semantic grouping for vision-language navigation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:14.452660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:14.452660Z digest=sha256:22b9cb46fdc944382ad5ad517677195d7bf73583958e4c3211fe7eedf1e40394

Observation 73b7b311-089d-4fc2-bd28-412dbbc65400 · outbound

This paper cites Etpnav: Evolving topological planning for vision-language navigation in continuous environments,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Etpnav: Evolving topological planning for vision-language navigation in continuous environments,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:14.669721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:14.669721Z digest=sha256:8393078cd74ee47d00047fd1f75665f314bb64720e664a77c0107de482c6b163

Observation 33b3d023-5eb4-4571-be07-19f7b42523a9 · outbound

This paper cites Dreamwalker: Mental planning for continuous vision- language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Dreamwalker: Mental planning for continuous vision- language navigation,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:14.811495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:14.811495Z digest=sha256:1fcabb58c4eeb9149ec6fcd62feaa148526f01e8b2dda9d952b7b83a6e908eb0

Observation 9280c9ef-0803-4978-a4bb-925df3a72173 · outbound

This paper cites Lookahead exploration with neural radiance representation for continuous vision-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Lookahead exploration with neural radiance representation for continuous vision-language navigation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:14.949609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:14.949609Z digest=sha256:863fadd6f5d084869a369054e5987f629879ee2bd28e84cee498ce22bd67d671

Observation e171ff47-c52f-4168-9b41-66264d09b133 · outbound

This paper cites Navgpt: Explicit reasoning in vision-and-language navigation with large language models,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Navgpt: Explicit reasoning in vision-and-language navigation with large language models,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.078012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.078012Z digest=sha256:f8f5dbebd6bd02640f5dcead1987dcc49489893d421a94afd682f2c87a346c1c

Observation d602a1c9-fd03-4ada-853f-3883742f3fae · outbound

This paper cites NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.215349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.215349Z digest=sha256:ddb1db7a6dd790b29f8b5d5e66a7906ee6949af7bffc4f4b88682a28524cba0c

Observation eb45f875-bba9-44ae-a15d-e832b434ccd8 · outbound

This paper cites Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.370486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.370486Z digest=sha256:d7c303e7290eb20c5078c207b6351a257e835d36251aa330fcadfa2eb341f2af

Observation 7b9be73c-fe2f-4bf4-b991-db29888e2848 · outbound

This paper cites VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation VLN-R1: Vision-Language Navigation via Reinforcement Fine-Tuning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.484911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.484911Z digest=sha256:8caa99996100b111d9875dd818bd9ee5cd571a0dda7f65d6df5042244a9744a2

Observation 2324cbe7-691c-44d6-8364-5693e57377f0 · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Dream to Control: Learning Behaviors by Latent Imagination

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.562072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.562072Z digest=sha256:31450dcc0272cb77c9585c3a0c518b30ab4e595ac2d87bfecacf841ae9889a78

Observation 4fcc13c2-3ca8-40ad-8f5f-589c8646421b · outbound

This paper cites Learning latent dynamics for planning from pixels,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Learning latent dynamics for planning from pixels,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.662143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.662143Z digest=sha256:95a799bc205015351b8c33706beaed639b5c9e6aac47b2d7dbab133a832c98ec

Observation 265013e2-a883-48fc-ad83-7dcdd5c5c1bb · outbound

This paper cites Pathdreamer: A world model for indoor navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Pathdreamer: A world model for indoor navigation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.764248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.764248Z digest=sha256:044e0dedfd6111cf4eed80b6e66c7a2dd4c363a0bbf56c618fcdcd58fe6007f9

Observation 86774cf2-987d-4613-9e14-a40661e47193 · outbound

This paper cites Panogen: Text-conditioned panoramic environment generation for vision-and- language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Panogen: Text-conditioned panoramic environment generation for vision-and- language navigation,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.859295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.859295Z digest=sha256:03a2154d24e9363a8abb15584d1040a395129b875f4864dc225938f5fe802b5a

Observation f991e654-ca05-425f-8ddf-aae835df3fe9 · outbound

This paper cites Do visual imaginations improve vision-and-language navigation agents?.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Do visual imaginations improve vision-and-language navigation agents?

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:15.934047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:15.934047Z digest=sha256:27d0482fa490ba40a422c6d48fb01b06dd34e0af5ad1c9129d2eed928be5b336

Observation 884692f4-cc5f-4464-add7-c58fb3609d2c · outbound

This paper cites Navmorph: A self-evolving world model for vision-and-language navigation in continuous environments,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Navmorph: A self-evolving world model for vision-and-language navigation in continuous environments,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.018233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.018233Z digest=sha256:9b3ba1ab31d7cd8bde701b0825bc5e3053f656462ab101f1686ce390c164b278

Observation d2477e33-35cd-4bd7-8736-a9bb632fb8b2 · outbound

This paper cites DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation DreamNav: A Trajectory-Based Imaginative Framework for Zero-Shot Vision-and-Language Navigation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.086989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.086989Z digest=sha256:6eb091cbab9051552336f9d408abe22b0177d7be583d182e695a3ba90f9cbefd

Observation 0d0b3b89-d582-4ff8-a181-719840c361dd · outbound

This paper cites Navigation world models,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Navigation world models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.153246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.153246Z digest=sha256:3e4416f1fb826dde4f52c693d9cb6c539190ccfabf3fbdb9c206cf59fdfaa2d7

Observation 68d25112-d192-46bd-9340-03aeb4a6d3e1 · outbound

This paper cites The expression of a tensor or a polyadic as a sum of products,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation The expression of a tensor or a polyadic as a sum of products,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.256185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.256185Z digest=sha256:e82a152b8848dd906ce28e2498b668562e413ffa1f0aa8a2f1aad986870071cc

Observation 169de0c5-32ff-4053-9a1c-023f9a84d16c · outbound

This paper cites Some mathematical notes on three-mode factor analysis,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Some mathematical notes on three-mode factor analysis,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.309801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.309801Z digest=sha256:639c2f3dd60a70a70f10af266e358a31182957f3f4f0f5622a73cf62a256daee

Observation 67c74862-7d94-4424-b667-0e3cfc00d87f · outbound

This paper cites Disentangling factors of variation in deep representations using adversarial training.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Disentangling factors of variation in deep representations using adversarial training

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.377030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.377030Z digest=sha256:b194759e77c9001575b54e0a0e77c7bbb1c76f891ce655232af489990759b603

Observation b8d29965-615d-40a2-a8b2-b62f2a2c5835 · outbound

This paper cites Hamiltonian latent operators for content and motion disentanglement in image sequences,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Hamiltonian latent operators for content and motion disentanglement in image sequences,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.447675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.447675Z digest=sha256:ac39316086fe39eac9ab63bd0104a95a58862fcc2c28a32201ff697190ee6256

Observation a25656f3-e7e4-47e0-8793-70c27f10d831 · outbound

This paper cites Iso-dream: Isolating and leveraging noncontrollable visual dynamics in world models,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Iso-dream: Isolating and leveraging noncontrollable visual dynamics in world models,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.568902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.568902Z digest=sha256:0f7067878a50435a30bee3fa67d5cb6c2c0763630edb16a70ed4e57f5ac8a6b6

Observation 0fcc3d26-ca53-40db-b1a1-3a05098615ca · outbound

This paper cites AdaWorld: Learning Adaptable World Models with Latent Actions.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation AdaWorld: Learning Adaptable World Models with Latent Actions

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.735012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.735012Z digest=sha256:1f92078181c79eb751271a0ee8adbdf4dc37d3dc16846bf4f2e6f8a053b2bae4

Observation f2bcc129-7ad7-4880-82c5-ed3c0c899b53 · outbound

This paper cites Factored Latent Action World Models.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Factored Latent Action World Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:16.890582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:16.890582Z digest=sha256:22bce5ec5db530be63cbb1e68f5acce88a3164775e427e36c6ba8b0f32a93eaa

Observation df0c69d3-710d-43f7-b214-086d5db17f22 · outbound

This paper cites Compressing neural networks using the variational information bottleneck,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Compressing neural networks using the variational information bottleneck,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.090795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.090795Z digest=sha256:d53200cfc1549c5ba6cf71fb9ea4523ade135020c78bd696ed03e7b98749310e

Observation 242e0bd2-5197-4123-9615-7d66c4892370 · outbound

This paper cites Learning Sparse Latent Representations with the Deep Copula Information Bottleneck.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Learning Sparse Latent Representations with the Deep Copula Information Bottleneck

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.229949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.229949Z digest=sha256:b749968f7bafd9926d279badd274b0858f4fe01c2f09dbcc8b089c7b95094897

Observation 2d91c6ef-cf2b-4dab-891f-340f62ead431 · outbound

This paper cites Attention-Based Guided Structured Sparsity of Deep Neural Networks.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Attention-Based Guided Structured Sparsity of Deep Neural Networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.387571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.387571Z digest=sha256:98042b779691e08be772e0f8667a67135c753cfaf3e1ab28fa650207c1de0830

Observation 6df6b89a-4594-4fa5-bdaa-e6038b864025 · outbound

This paper cites Learning latent dynamic robust representations for world models,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Learning latent dynamic robust representations for world models,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.503590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.503590Z digest=sha256:5457f2acfe14467462051f8f2f30e470c3370030e95449e7a33e3852c606d42d

Observation 6b4af7ba-75ed-4da6-a22d-74e9fcf1a821 · outbound

This paper cites Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.843603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.843603Z digest=sha256:8aae35ee3b91b97084ff595333a91599f31eadaa7576a0c4a0fbe22b970b7b8d

Observation c518e69c-dbe0-4402-a81f-c67ff2897e50 · outbound

This paper cites Enhanced structured lasso pruning with class-wise information,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Enhanced structured lasso pruning with class-wise information,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.954779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.954779Z digest=sha256:dda4535dc1a66ad1802faea55829a4a18f2e8ca8e83cd5e06825301e0a4e860a

Observation 4281c3b9-4974-4963-b353-57acb2594cae · outbound

This paper cites Object-centric learning with slot attention,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Object-centric learning with slot attention,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.070503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.070503Z digest=sha256:42112a58872e4292799584be10c74590e13728903d4bb2a4ca785fc220b71651

Observation 47d2cd77-f1d9-4740-bb0b-b1424a25cd08 · outbound

This paper cites Qwen2.5 Technical Report.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Qwen2.5 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.152576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.152576Z digest=sha256:cb8437af870337213fb3e14d24abef02c80f137ede062b2e54013d0654c7d261

Observation 70c3d765-cbc9-4994-a541-b81e096cfdd1 · outbound

This paper cites Tensor decompositions and applications,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Tensor decompositions and applications,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.248349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.248349Z digest=sha256:90423d543903cdc694d67dff3873df16524a0194891b9f691b3de83c734c8b0d

Observation 5d1612b5-3e74-4084-9183-e3d8f1b719a5 · outbound

This paper cites Tensorizing neural networks,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Tensorizing neural networks,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.353740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.353740Z digest=sha256:649e536ac8cc8c29894cd466c0f304e9d5ac0d438b0d8a11dc95a915b3967473

Observation c36c752a-67d2-42d8-8463-0ec33cb781b9 · outbound

This paper cites Deep Variational Information Bottleneck.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Deep Variational Information Bottleneck

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.465247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.465247Z digest=sha256:2c638ce426f6982d4f9720bce411e88b239a6b0d5b79cb31da05bb3b96914a24

Observation 9b199ddd-48a7-4b9a-b89d-45cdab281688 · outbound

This paper cites The information bottleneck method.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation The information bottleneck method

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.571987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.571987Z digest=sha256:8a745e88be6d3b263c0a52de33c1f5e68c5d232c3478242c0663752d6f9ba428

Observation 4d2e72bf-55e3-4d27-a68f-5ce0d0b2cd87 · outbound

This paper cites Sim-2-sim transfer for vision-and-language navigation in continuous environments,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Sim-2-sim transfer for vision-and-language navigation in continuous environments,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.690277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.690277Z digest=sha256:253c8e82c5d67b946ce32bfb4d41d9d58b453f1f1978e4161ab0d147817b213c

Observation 6d708582-f21a-4849-891c-655cc7a777e9 · outbound

This paper cites Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.769727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.769727Z digest=sha256:4d1d6b27595a2dded7ac44aef1005dd330bb40903bf246cb910bf2b57d90aada

Observation 6399b1a6-e056-4e40-b7f2-b4d8d0c40d10 · outbound

This paper cites Dreamwalker: Mental planning for continuous vision- language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Dreamwalker: Mental planning for continuous vision- language navigation,

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.853036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.853036Z digest=sha256:b2fa6f19453ebf4f4b2a3da255e11de74c5ad1c7f7d17c939b283590409e0a75

Observation b9c32144-3293-4884-9531-3e56639c8a79 · outbound

This paper cites Gridmm: Grid memory map for vision-and-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Gridmm: Grid memory map for vision-and-language navigation,

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:18.937944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:18.937944Z digest=sha256:32e3966cd4d21c5dd7e97fb1852e1f6c3f8b856ece2209f3c5673e022d2ce54d

Observation 0899cc28-37fb-4e2e-b629-ba12eccde1ed · outbound

This paper cites BEVBert: Multimodal Map Pre-training for Language-guided Navigation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.022152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.022152Z digest=sha256:cc086b0461240aa107746788e0ad0532aed9c195f9037f0cad80c8bb4c553485

Observation 6c6208cd-036d-4790-8d12-6b73e3219350 · outbound

This paper cites Fast-slow test-time adaptation for online vision-and-language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Fast-slow test-time adaptation for online vision-and-language navigation,

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.104462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.104462Z digest=sha256:6fa25a2fde9a777b1e6ee167bdc2e96afa17d1a75377e40eaa73fef2953ec130

Observation 64ae8c9c-df64-4ef5-8433-a9de0d576f88 · outbound

This paper cites Language-aligned waypoint (law) supervi- sion for vision-and-language navigation in continuous environments,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Language-aligned waypoint (law) supervi- sion for vision-and-language navigation in continuous environments,

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.174228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.174228Z digest=sha256:a55da3e82279bc031f1812ac3fbf04d9690e55d8f11213926e459691915bdcd4

Observation 06fad732-dd8d-492f-8382-863321962362 · outbound

This paper cites Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Affordances-Oriented Planning using Foundation Models for Continuous Vision-Language Navigation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.178700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.178700Z digest=sha256:f4bccdbb5be5fba0ef084a579e1562960206eaf2982c459fc91909fa86aaac77

Observation ed6979da-f556-412c-9d42-8254340c37a0 · outbound

This paper cites 1st Place Solutions for RxR-Habitat Vision-and-Language Navigation Competition (CVPR 2022).

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation 1st Place Solutions for RxR-Habitat Vision-and-Language Navigation Competition (CVPR 2022)

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.181601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.181601Z digest=sha256:edee2bdafb0d973ba145a69a2409f2bdae913e1dba18749572ea095c5dbc16ad

Observation 4b85743e-5f66-48a6-add6-e28ff659a34b · outbound

This paper cites General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation General Evaluation for Instruction Conditioned Navigation using Dynamic Time Warping

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.184355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.184355Z digest=sha256:19bd097a98a755807c83471cf060feddce90f8631278f2d5c3974e168316e5d0

Observation 47361b57-e13d-4e66-a05f-dfa6cbe76118 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.187322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.187322Z digest=sha256:fe350863c2a86e6ce763caf0930361565addf43732a2c80c40b1102ffcf58ec7

Observation da64fecb-2b57-48df-9eec-ec34b719c643 · outbound

This paper cites Deep residual learning for image recognition,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Deep residual learning for image recognition,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.190214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.190214Z digest=sha256:3216931870ecf000a7d6a184528a60f4f807b765d7b5832e1b18ab28ddbf3d89

Observation b29a330f-ffce-496d-8f11-9737d92265cb · outbound

This paper cites Cross- modal map learning for vision and language navigation,.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Cross- modal map learning for vision and language navigation,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.193062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.193062Z digest=sha256:05610b11ed7cd1009d1fae0628db87b5e6068751e9be42688bf60919954da370

Observation d18b334d-af2e-4686-b0cb-9dd7153719b8 · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.195676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.195676Z digest=sha256:f694cc2d31bec45e4955b7b3f6eb6807dec26809fbd53f70586a027c0ea3dc19

Observation 9445d9d9-0ecb-4833-b8f5-612d81dfb01d · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.198670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.198670Z digest=sha256:5a332da3063c7187457c5c82e42b8b82ac17fe7c160bb6daf71b9748750677f9

Observation c87af2bf-3a10-4f05-83d5-2e3ae50a5259 · outbound

This paper cites Learning Latent Dynamic Robust Representations for World Models.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation Learning Latent Dynamic Robust Representations for World Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:17.714001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:17.714001Z digest=sha256:9410e6ab738c13de2b69d2fed3d38b96ac453d27ad2c6aea800d158ab4afe941

Pith citing papers

No inbound Pith citation observations are available.