Pith. sign in

Paper Citation Record · LEDGER

BEVBert: Multimodal Map Pre-training for Language-guided Navigation

As of 24 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2212.04385.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.04385 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:06:07.499847Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:26:54.066214Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ebe58894-a4b8-4119-b457-6a51c6b8dfba · inbound

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation cites this paper.

NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:55:20.409420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T04:55:20.362512Z digest=sha256:de95177420fbd26373162881460cc0204d5435872d81ade04835e7e1bd3c8c8e

Observation 02cfc938-f6e9-4897-98b1-d6ee4feea425 · inbound

NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation cites this paper.

NavAgent: Multi-scale Urban Street View Fusion For UAV Embodied Vision-and-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T21:36:19.661906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T21:36:19.661906Z digest=sha256:b3338969fe5bc47ed7f7eaf289116ebbfaab91f4db467e7e3c7454da15872905

Observation 10502d36-f334-44f4-b8a3-f7ef32824ca1 · inbound

SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts cites this paper.

SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T20:40:39.376383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:40:39.376383Z digest=sha256:908400b70d5a692e11e529246a5631b859b940bb0c805a5e5bdb4caba5e1fecf

Observation d229a878-0958-4f5e-ba05-64aa8e326b0c · inbound

Think Hierarchically, Act Dynamically: Hierarchical Multi-modal Fusion and Reasoning for Vision-and-Language Navigation cites this paper.

Think Hierarchically, Act Dynamically: Hierarchical Multi-modal Fusion and Reasoning for Vision-and-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:06:07.499847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:06:07.499847Z digest=sha256:5e35e55bc5d7030feba54ada50cf65707c913466671fc158434faef4992bb2bf

Observation 6c700b45-057a-4f36-b138-e8092009fdce · inbound

CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation cites this paper.

CoNav: Collaborative Cross-Modal Reasoning for Embodied Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:51.260957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:51.260957Z digest=sha256:05539e838723417c4c4cc6baeb4bad6c889b4b19dcfbdefeddeb4f463cc84f81

Observation 8515c56d-e73d-4b7b-be09-c17388a5d911 · inbound

Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation cites this paper.

Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:52:03.792370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:52:03.792370Z digest=sha256:33a091b7b236d425109924c7acddab76d00346a4739f69b1cb5c595057b86b84

Observation 8d8a8985-778e-49c2-b5c0-3a0e79053f48 · inbound

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation cites this paper.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.577923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.577923Z digest=sha256:2c5eaaef67646c2193bb541d56a455a2dfd5976cd57358cd53cca472d15c70da

Observation 6844dbe5-9679-49a8-bc10-f4558754dd79 · inbound

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments cites this paper.

NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:48:22.838916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:48:22.838916Z digest=sha256:3f4548c2ff9f1a8fccd6420a44ad2e25742f40673daaeec57f8110ac3d98f520

Observation 701d898e-fdc6-4ed5-9054-0438f6046e46 · inbound

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation cites this paper.

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:05:15.982288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-17T21:02:38.013115Z digest=sha256:63589bdbb67934c0566509c519959c01734a887f580d69ed4ab6edfd0f6a23ed

Observation cb1db964-b710-47fe-b6d4-6105fbb5ae35 · inbound

HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System cites this paper.

HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T18:10:13.897082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:10:13.897082Z digest=sha256:5e405475bf8f5a89ccd9b200490887102771b338faedd4a0395987f567604015

Observation 7665012f-0536-40e8-8e37-b36cd17bb917 · inbound

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs cites this paper.

LightZeroNav: Zero-Shot Vision Language Navigation in Continuous Environments Based on Lightweight VLMs BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:44:07.845806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-21T10:41:47.297072Z digest=sha256:a39a23ecab17062a5e9e0c0b40f2ac9fb36159b976994f405ad3ae93677738c8

Observation 0fade2b9-56f9-4a89-84ed-b91860c0a2d5 · inbound

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions cites this paper.

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:53:04.305421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T08:49:33.107658Z digest=sha256:ea4244787642d1b2fde015b714281c3c115619263b061c5865917ca4a9374c57

Observation 403e845d-de07-482b-894d-78bf9f1d6f2d · inbound

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation cites this paper.

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:51:45.893511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-10T06:50:34.310831Z digest=sha256:e6cef5fea2e68d795ded2c2f30233e2640214416286acddc7eb6d625f6207817

Observation 6ff4f4a3-6243-4179-bf13-6ef065173d82 · inbound

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation cites this paper.

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:06:14.900560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T02:04:57.627160Z digest=sha256:025954d6da60a494da468042e149ee32ae2d9cfefe7e4020582e50b1c1f39b28

Observation 989c518d-6a76-4494-8fa5-8735a378b2e9 · inbound

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation cites this paper.

GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:46:10.657852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-22T06:45:35.920991Z digest=sha256:7ed55b7661eb5e4b0430e149907e8c8dd88ebcb9de5ea36e9ec85852697999f1

Observation f4749ee8-2d9f-46e1-be8a-9d0851fd3236 · inbound

Large Depth Completion Model from Sparse Observations cites this paper.

Large Depth Completion Model from Sparse Observations BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:23:15.260029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T08:19:01.715788Z digest=sha256:4b46affa9bd4bead0f8e969346f80e4003f69535b6d0bad1b27f1f24ef0425e9

Observation 7edd71a2-9bbb-4ff3-809e-b820123e3123 · inbound

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation cites this paper.

Ask When It Pays: Cost-Aware Open-Ended Interaction for Instance Goal Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:26:26.147588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-28T11:01:08.668612Z digest=sha256:0f455a6bb410fde48dd95ba6781bddd8af868fef1397785c4863afa985fcbcee

Observation 0561c1c0-5f30-4094-854e-45aaba081f13 · inbound

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks cites this paper.

SpikeVLA: Vision-Language-Action Models with Spiking Neural Networks BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:33:54.610574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-29T04:42:23.040915Z digest=sha256:7e1fd8202b45cd852aa1bd5178fe9ace6ca480a6617a009b267947a55c3c8699

Observation 56c6f726-63d1-4d18-99e4-d6a3e14cc72d · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:26:54.068087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-02T11:17:26.529397Z digest=sha256:8b6ef4a53a6372c4fa9e27753fe4525b81e85424503856369374d9024f224407

Observation f894773b-7996-42ee-ba10-f32dc45ef0f1 · inbound

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation cites this paper.

DART-VLN: Test-Time Memory Decay and Anti-Loop Regularization for Discrete Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:40:14.040530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:40:14.040530Z digest=sha256:49f7788f0b054e6416bbb24c0cc5a88bd2083f7a578a116aab7adf1f843ca1e3

Observation 0899cc28-37fb-4e2e-b629-ba12eccde1ed · inbound

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation cites this paper.

Towards Dual-Brain Minimal Sufficient Representation for Vision-Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T03:26:19.022152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:26:19.022152Z digest=sha256:1b0b0e569300c8bebe4d9ae77ce8da91b473c94df31a9253452160fddb0b68ab

Observation 1b1b750b-06c1-43b7-8017-1a3d444161e0 · inbound

Goal-oriented Navigation Instruction Generation with Tour Video Priors cites this paper.

Goal-oriented Navigation Instruction Generation with Tour Video Priors BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:34:34.982299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:34:34.982299Z digest=sha256:80180cc138d998b117d541e957f536c3e9215b2a10297c53579796b0ba1f37e8

Observation 07e40045-e9f5-446b-b91e-6ba961d7f551 · inbound

DaViNCi: A Dataset Towards Outdoor Vision-and-Language Navigation with Continuous Actions and Dynamic Elements cites this paper.

DaViNCi: A Dataset Towards Outdoor Vision-and-Language Navigation with Continuous Actions and Dynamic Elements BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T00:27:57.368415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:27:57.368415Z digest=sha256:03083a555125fc012e4366f8de39cb9d042783dcc3bfb271ecd9194907e79429

Observation 1c31f4c2-5388-419f-b3b4-9f4201742a4d · inbound

AirForesight: Current-to-Future Spatial Map Imagination with Cross-Space Planning Consistency for UAV-VLN cites this paper.

AirForesight: Current-to-Future Spatial Map Imagination with Cross-Space Planning Consistency for UAV-VLN BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:23:27.650526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:23:27.650526Z digest=sha256:5582e6960c85fca45176b27944a1e70a2320317b387c91ca1c9635001b2b987f