Pith. sign in

Paper Citation Record · LEDGER

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation

As of 23 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2506.15757.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15757 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:39:25.788223Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T02:24:42.530103Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T02:27:07.269979Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bd1cc1ea-f511-4b8d-a2e3-bae48f3a28ed · outbound

This paper cites Vision-and-Language Navigation: A Survey of Tasks, Methods, and Future Directions.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vision-and-Language Navigation: A Survey of Tasks, Methods, and Future Directions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.453113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.453113Z digest=sha256:69d7d039068b0902fbbc7a91f4024de864a9a958b4d8b08f787183ca0c6dafcc

Observation 9fb74853-708f-4bc5-9090-6ed43ad5011f · outbound

This paper cites Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.459368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.459368Z digest=sha256:916731496ff4c7c31560286f41cdcdede5295cbc25d65f13814e5f8004873378

Observation 3283d03f-cd82-4b05-85c4-d9c9fe97cd27 · outbound

This paper cites Speaker- follower models for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Speaker- follower models for vision-and-language navigation,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.464502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.464502Z digest=sha256:f4d63c41a9803e7930ae49c57f632c5476137166c34c46df11e58c8b40cedf98

Observation a90c3d0b-98c3-4be1-af88-35e295abafbd · outbound

This paper cites Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.469518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.469518Z digest=sha256:2040e7a157ed6826b34035d222d4d07436d2211a7549beec9321dfcefbf975fd

Observation 5be27518-1769-4b2a-a253-26a2b09badb5 · outbound

This paper cites Towards learning a generic agent for vision-and-language navigation via pre-training,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Towards learning a generic agent for vision-and-language navigation via pre-training,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.474964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.474964Z digest=sha256:08020c9116a3fe6711f7faf0a20d0141942539f59f4e9b0ea163049bc962fc68

Observation de0d2dd0-2432-4cda-944e-aac17b702f30 · outbound

This paper cites A Recurrent Vision-and-Language BERT for Navigation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation A Recurrent Vision-and-Language BERT for Navigation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.481181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.481181Z digest=sha256:9cbfed3e1489a935e9ce723f949b793d1fa079d410545395b31fe52df5f04f22

Observation 38252b74-377d-4100-b5eb-71ed85eb92ba · outbound

This paper cites Scaling data generation in vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Scaling data generation in vision-and-language navigation,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.487051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.487051Z digest=sha256:a7f76d8aea9e67d06cee27fc206f40eaf02d8738258f6938a4d271fc90347d4a

Observation ebb01259-d9bc-4ad1-8c51-d1db20632849 · outbound

This paper cites Navgpt: Explicit reasoning in vision- and-language navigation with large language models,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Navgpt: Explicit reasoning in vision- and-language navigation with large language models,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.492007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.492007Z digest=sha256:455dee71c74515b3e092cadd843e916cfed69802a9c5ca086e542df585d28b88

Observation fdc28d61-05c1-44a5-88de-c9ceee604ee0 · outbound

This paper cites MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation MapGPT: Map-Guided Prompting with Adaptive Path Planning for Vision-and-Language Navigation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.496995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.496995Z digest=sha256:275fd96979464c94c27c75071d5c28ace37db5d4c3344283cfb06ae2a1722762

Observation b019194b-31f4-47a3-af8f-572b447ab7ae · outbound

This paper cites Discuss before moving: Visual language navigation via multi-expert discussions,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Discuss before moving: Visual language navigation via multi-expert discussions,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.502430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.502430Z digest=sha256:e259487300535e9e9ef54ebf19b23a48991a354e7b6b5bd170efdc0fe99de553

Observation e3742a15-024f-46e7-ad89-d2e8f592eae5 · outbound

This paper cites NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.507096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.507096Z digest=sha256:fc6582ceae96d31d7fb144ce4f322788ebabb3de7e01f1927c82af7ed773f5f1

Observation 507036df-8e69-4be6-92b1-27c16618115e · outbound

This paper cites LangNav: Language as a Perceptual Representation for Navigation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation LangNav: Language as a Perceptual Representation for Navigation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.512113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.512113Z digest=sha256:a6f9712c755cacd579e40b052f8cc9ea1d964cfcd22dc5d8336ff569aa0afda8

Observation 758fdd73-cdd1-4188-8083-73bd85d9b1ef · outbound

This paper cites Towards learning a generalist model for embodied navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Towards learning a generalist model for embodied navigation,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.517864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.517864Z digest=sha256:0c3408ab69d7a547e5f5ad99aaf2bb460f95a42cf80440e1bcecd4de26ea71d1

Observation 71e0b4d0-c630-4972-9b65-c893607cf39a · outbound

This paper cites Deep residual learning for image recognition,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Deep residual learning for image recognition,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.522848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.522848Z digest=sha256:64c7c98a2c807c9cf91696885591bebecee557b63248fce4f8f71552fcbad34a

Observation 2e4f6656-3908-4555-82af-23561536e47e · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Learning transferable visual models from natural language supervision,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.527605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.527605Z digest=sha256:d67a4b92b831175d5f1fb18aa2814875f7a0d82456b91c74f1a3a1be48595f9d

Observation 3f84d61b-26bc-4c8a-9983-aa00808fcac0 · outbound

This paper cites Active open-vocabulary recognition: Let intelligent moving mitigate clip limitations,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Active open-vocabulary recognition: Let intelligent moving mitigate clip limitations,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.680434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.532464Z digest=sha256:6abda61edbb494e1dbcd995671445b38f0f500bf94087e7dd089bef487d0a13f

Observation 4d758a50-21a1-4a25-865c-cd578d5e443f · outbound

This paper cites Contrastive learning inverts the data generating process,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Contrastive learning inverts the data generating process,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.663338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.537664Z digest=sha256:d8613f96227fe7f2bf266b649da078a581f0765f367a19419148b34ea36dc9ea

Observation c69f5006-aac3-4f12-9cab-3289377b2144 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation A simple framework for contrastive learning of visual representations,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.542415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.542415Z digest=sha256:8ac90ed58c2e76bc3d5554435bb26987703fdf9813876758cbc1625c891dbc1a

Observation a9464707-2d21-4b26-8bb0-2ed6e775d1b9 · outbound

This paper cites Improved Baselines with Momentum Contrastive Learning.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Improved Baselines with Momentum Contrastive Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.547722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.547722Z digest=sha256:285b3fc8839739cf9790e42cc585db8cb1009519ce02182129c4387e6000f15f

Observation 4c192bee-737a-4728-8c51-1e295718fbd1 · outbound

This paper cites Visual instruction tuning,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Visual instruction tuning,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.553663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.553663Z digest=sha256:576208b381ba166ff3c9f1e14d96e0c099a3d9688478550b2eb0f893d3bc5f03

Observation 284d2286-bd09-4ef0-a195-03b595f49ecc · outbound

This paper cites Simple but effective: Clip embeddings for embodied ai,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Simple but effective: Clip embeddings for embodied ai,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.558603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.558603Z digest=sha256:c454e5ae1d7d0d6711af01fa8e07a17e6dd91a7e0adbbf088f62a2c99e21b525

Observation aaa87f89-a0a4-4483-bf02-ec75a152bebb · outbound

This paper cites Think global, act local: Dual-scale graph transformer for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Think global, act local: Dual-scale graph transformer for vision-and-language navigation,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.563499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.563499Z digest=sha256:a22609228af1ffee9ba2c1b09f484a7516d9a4c9400b81cd20ffc9a1585d48d7

Observation a85404ef-e6e3-4c4e-8c2f-6df0788ef6cf · outbound

This paper cites History aware multimodal transformer for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation History aware multimodal transformer for vision-and-language navigation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.568151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.568151Z digest=sha256:85aca83ba7096ea2d2b816cf136941ef0e83dcfa799a0043c0deb2fc469ce45b

Observation d1af1162-ca36-4f61-818d-3fbcc9b2a92d · outbound

This paper cites Scene-intuitive agent for remote embodied visual grounding,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Scene-intuitive agent for remote embodied visual grounding,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.594278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.573140Z digest=sha256:40d32ebe4dd5e6415bcbc770c3208e3405e0e38a9df858a7d9d971574f3cf681

Observation 8d8a8985-778e-49c2-b5c0-3a0e79053f48 · outbound

This paper cites BEVBert: Multimodal Map Pre-training for Language-guided Navigation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation BEVBert: Multimodal Map Pre-training for Language-guided Navigation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.577923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.577923Z digest=sha256:2c5eaaef67646c2193bb541d56a455a2dfd5976cd57358cd53cca472d15c70da

Observation 9390174e-37f3-4d9f-a4f6-aa9918756524 · outbound

This paper cites Vision- and-language navigation via causal learning,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vision- and-language navigation via causal learning,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.577405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.583816Z digest=sha256:e6ce3797a2c2e9cf46a77c580d74a46b00bbe9437755df35601ca0ceb39cd531

Observation f0eca0c7-7049-4b8b-b5e9-caad00bf28e1 · outbound

This paper cites Gridmm: Grid memory map for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Gridmm: Grid memory map for vision-and-language navigation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.588635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.588635Z digest=sha256:d19a80fdcca36c5efb4d412a5bb48fb7280bbfb805e36033ba99c1e82a030f53

Observation c1afeb9f-75d0-4c92-8e30-265c5baf4b0d · outbound

This paper cites NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation NavGPT-2: Unleashing Navigational Reasoning Capability for Large Vision-Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.593380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.593380Z digest=sha256:e2eff9d3747a7ac23879c1044ef5fe31dca0e926276fd069429120a7f2557f3c

Observation 9950bb5f-f4d4-4b88-ac1c-b7007782de0b · outbound

This paper cites Reverie: Remote embodied visual referring expression in real indoor environments,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Reverie: Remote embodied visual referring expression in real indoor environments,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.599427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.599427Z digest=sha256:2153cf8856200cb9db45004f9b80cc5cc7c3c9cd3e31f28b1c0c8e76e4f96c79

Observation 5ff0db1a-1258-4e52-8a9c-93472af3a052 · outbound

This paper cites Soon: Scenario oriented object navigation with graph-based exploration,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Soon: Scenario oriented object navigation with graph-based exploration,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.538984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.604190Z digest=sha256:1bfa53f0c248b2f62102184e6495814fc72d8bd09087a3238f47a8bea35e5eb4

Observation 984b1c40-0aba-4e0d-8143-5b5127dd203f · outbound

This paper cites Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.608962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.608962Z digest=sha256:d0074b05ec80200820385f5d6c0236049d172225814a13b7f3831de29cda355a

Observation bc278a23-7fe6-412f-a0bd-6facd4b77ca8 · outbound

This paper cites MiniVLN: Efficient Vision-and-Language Navigation by Progressive Knowledge Distillation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation MiniVLN: Efficient Vision-and-Language Navigation by Progressive Knowledge Distillation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.613611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.613611Z digest=sha256:c45469c5fe64f60901825da26923a0785bd405128b57149b7d960f0616d2e250

Observation 91825b55-51f6-488e-9177-511db69dd192 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation LLaMA: Open and Efficient Foundation Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.619445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.619445Z digest=sha256:0d6e6154fb8580023dbb6019b70e3f6d4b70c2eb6d13bf11f2a593569a799a64

Observation 937dfe1b-178f-4aad-b6e8-28a15ab99a54 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.624381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.624381Z digest=sha256:26a6bee7b56e525f55fe5773a3e664dab2021a2b9320b51c94f89fbf6e901268

Observation beb4c570-c3a1-416a-bc65-944767e4d918 · outbound

This paper cites Scaling instruction-finetuned language models,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Scaling instruction-finetuned language models,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.501461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.629496Z digest=sha256:1996f4f90c98b4e93458c4823b86f5923aa8e5256e8e4ea4af61bde37b4c0c6f

Observation d0e18d11-0828-4841-807c-8a58c4275979 · outbound

This paper cites GPT-4 Technical Report.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation GPT-4 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.634062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.634062Z digest=sha256:0091f89c8908ef77ad7f676de9897b7e31ecffb7facc4139f07dc97c0e82352f

Observation 682fdc34-c0f0-4493-8e6d-df7215d56bb8 · outbound

This paper cites On Evaluation of Embodied Navigation Agents.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation On Evaluation of Embodied Navigation Agents

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.639104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.639104Z digest=sha256:534d4f98fda14e99a073f9e0647adcb1e3d3ca522fdb4d638baf4f44ad76fa0e

Observation 764a5fda-0b91-407f-bfed-45f230ff96ba · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.644022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.644022Z digest=sha256:eb9eebfbe219bedbe7963b14658394736179aba77187f07ca111a8715b24fe86

Observation 9cbe6b65-9f58-4199-ab23-4adb5844bb2d · outbound

This paper cites Airbert: In-domain pretraining for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Airbert: In-domain pretraining for vision-and-language navigation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.482429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.649158Z digest=sha256:56301beb3d924415f1bef47a5e91005137974d52d0909af6c5dab960530e3757

Observation 42d20c5e-17fd-4e72-aff4-c55f78beeee6 · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Llava-next: Improved reasoning, ocr, and world knowledge,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.654197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.654197Z digest=sha256:7791ad7345bea779062215cc01275324089f6065da23e93df0c750999a55eac2

Observation bdf64db4-ca48-4047-b024-f3386bf9b2a0 · outbound

This paper cites OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.659043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.659043Z digest=sha256:6783325eeff9db402276cb944d982d831f5a851a7a2b7e690501b095948c1f31

Observation 3ebd3d9f-80c3-4fcb-9877-4d49cd3dd2ab · outbound

This paper cites Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.664140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.664140Z digest=sha256:a5f47b1cef122fc68249cb0907a2060fb3b6671eaccf90cdece661bdeb9b608f

Observation 74e3ea3d-b62b-48c8-8655-1ab39d1c24b0 · outbound

This paper cites Vision-language models for vision tasks: A survey,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vision-language models for vision tasks: A survey,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.669880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.669880Z digest=sha256:36bf84959a7d84f16446864fb0efc9f8f9166648cd02d7d66605f73dec0a1b79

Observation 1ee5930a-e6c3-4b35-ab2f-b76b51194f7e · outbound

This paper cites Vinl: Visual navigation and locomotion over obstacles,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vinl: Visual navigation and locomotion over obstacles,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.430518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.675678Z digest=sha256:2612d21aef61d659438696763b4d7d115df49bc00ae9e3f09afaab3899137b97

Observation f80c1e5b-d7a2-46a6-b820-1433cbce67ea · outbound

This paper cites Exploitation- guided exploration for semantic embodied navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Exploitation- guided exploration for semantic embodied navigation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.413011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.680905Z digest=sha256:e039f7a9c59f5cb4eab6f5b0aaa810a571c02b5cfbe202d9e9703272e8bd0408

Observation 7d795514-ae1e-474d-8198-c5d3c4ec0dbc · outbound

This paper cites The regretful agent: Heuristic-aided navigation through progress estimation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation The regretful agent: Heuristic-aided navigation through progress estimation,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.395039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.685796Z digest=sha256:aa32e26634c8bdf4db1de32baf7f6c86fb5bcb72ceb4249a9b9e4009eeba56b4

Observation 4bc5e4de-33f8-470e-aee5-78fb574f2112 · outbound

This paper cites Tactical rewind: Self-correction via backtracking in vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Tactical rewind: Self-correction via backtracking in vision-and-language navigation,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.376857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.690597Z digest=sha256:b992e2181b99b61186af439cfc69497ee6e0fd74835f72f6a0f0ae79575f385e

Observation 8a9c334a-c2a3-4e30-8ab9-04cd5b0dd13b · outbound

This paper cites Bird’s-eye-view scene graph for vision-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Bird’s-eye-view scene graph for vision-language navigation,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.359153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.695501Z digest=sha256:34fbb3c3dd29213596500368904a815dbd9e883fee041059cbc96b7f08deee75

Observation 084a7f62-76d9-4c58-99ca-7a8922364a80 · outbound

This paper cites Visual language maps for robot navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Visual language maps for robot navigation,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.700838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.700838Z digest=sha256:21f14ab4131fe2b4957ce403da065d5fa9895ba69382b1bf9acb96f5d30730d0

Observation 88515a3b-a629-46c6-8201-c15d708d48a8 · outbound

This paper cites Robot navigation in unseen envi- ronments using coarse maps,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Robot navigation in unseen envi- ronments using coarse maps,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.328602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.705663Z digest=sha256:72989aa62d3701c08ca583b092db6c769eee46276e9fa974043bb5da4800aa06

Observation 28cf36ac-c3c6-410a-8969-7bf9c923697c · outbound

This paper cites Placenav: Topological navigation through place recognition,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Placenav: Topological navigation through place recognition,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.310669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.710743Z digest=sha256:6f853339221e673b4e59dfa497241c6d0d7ac7aa6deec05e3353a1a230af8288

Observation 6b492019-af6d-45dd-8de0-4ae6ae72c426 · outbound

This paper cites Kerm: Knowledge enhanced reasoning for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Kerm: Knowledge enhanced reasoning for vision-and-language navigation,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.716682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.716682Z digest=sha256:6e99dbd0f9a274fd80eb613df2f95422bba5ab6a0a179f08516beb6f1d59c9f6

Observation 125390ea-c04e-448c-ab91-1a5aee40d2f4 · outbound

This paper cites Zero-shot object goal visual navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Zero-shot object goal visual navigation,

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.721469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.721469Z digest=sha256:1e523af439d94f280ca32ebdb08f68feb4c5c1a3b9a4c8e7a16a4e13544a5c3e

Observation 66672cb0-fc31-44dd-934c-f125d139129e · outbound

This paper cites Wavn: Wide area visual navigation for large-scale, gps-denied environments,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Wavn: Wide area visual navigation for large-scale, gps-denied environments,

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.271173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.726294Z digest=sha256:53b77aa74006a9ce994267fe53c49f66495de0478157e9897c1ee801ba6bbc1e

Observation c01d2c48-6b09-4fde-8e44-e90280a9dc0c · outbound

This paper cites LOC-ZSON: Language-driven Object-Centric Zero-Shot Object Retrieval and Navigation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation LOC-ZSON: Language-driven Object-Centric Zero-Shot Object Retrieval and Navigation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.731371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.731371Z digest=sha256:cdb3125d58ee12e20d411650c9f5402c95bea52a325abcf12a2b12941939bc0d

Observation b0454822-8405-444d-84a4-8fd3faf95c37 · outbound

This paper cites Aligning Knowledge Graph with Visual Perception for Object-goal Navigation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Aligning Knowledge Graph with Visual Perception for Object-goal Navigation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.736583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.736583Z digest=sha256:f839b46304723a23ffd3ea8788ec53fdaa6a1642dd02d1439917679187a6d6e1

Observation 3e696c35-da0d-4eb9-911d-21a87d7836c8 · outbound

This paper cites Improving vision-and-language navigation with image-text pairs from the web,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Improving vision-and-language navigation with image-text pairs from the web,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.253355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.741309Z digest=sha256:cead8b58db80dd5162eefefb094b983b8b1fe2db9fee28545bd6f1fdccbad269

Observation 5ce1c420-210a-4589-bc99-7b874b528aa8 · outbound

This paper cites an unresolved cited work.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-15T19:39:26.237771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.746104Z digest=sha256:2ecbb8cd50858c191e3337f680cd4ac644ea6b1fd8de5e1d8e9e3722e16c7272

Observation 6095d16c-52d6-41e8-9699-e6ca31d5908b · outbound

This paper cites Panogen: Text-conditioned panoramic envi- ronment generation for vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Panogen: Text-conditioned panoramic envi- ronment generation for vision-and-language navigation,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.221288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.751931Z digest=sha256:7b60bb51f6b78e250ad124c8d34dd4e258840833670e56c736d5dbe21d5181a6

Observation d6d04429-fd8b-4606-9dc5-d08241435a5c · outbound

This paper cites Counterfactual vision-and-language navigation: Unravelling the unseen,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Counterfactual vision-and-language navigation: Unravelling the unseen,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.205031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.757074Z digest=sha256:0476ce1ca5cb9a61f6e58de139241087b37b456b3b16e2058bf58730d7eddbf1

Observation d2e8d691-acc5-4352-b158-3c28c6e97278 · outbound

This paper cites Counterfac- tual cycle-consistent learning for instruction following and generation in vision-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Counterfac- tual cycle-consistent learning for instruction following and generation in vision-language navigation,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.188406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.762112Z digest=sha256:f71992b0ab6108cc13f9e24fb3d1383be2a8370c857411ad43fc56537d570d84

Observation b82a011b-d363-4b20-8ad1-5657f294ba23 · outbound

This paper cites Transferable representation learning in vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Transferable representation learning in vision-and-language navigation,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.766940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.766940Z digest=sha256:a1b99ed35e210345845dfc975f14d12937ef7740d7acdd7e6ac07e960147a70b

Observation 7c8ab4f9-d387-478d-acfe-28bba2aaefa3 · outbound

This paper cites Self-Monitoring Navigation Agent via Auxiliary Progress Estimation.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Self-Monitoring Navigation Agent via Auxiliary Progress Estimation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.773063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.773063Z digest=sha256:97bb4dd4ff994a75d00aaa452fa26fe9ed2d0ad0132b42e95682f2bb26947a69

Observation 5f262a8c-ae1f-4737-ac0d-df159ef1301e · outbound

This paper cites Look before you leap: Bridging model-free and model-based reinforcement learning for planned-ahead vision-and-language navigation,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Look before you leap: Bridging model-free and model-based reinforcement learning for planned-ahead vision-and-language navigation,

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.778043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.778043Z digest=sha256:e2458677a94ea61b7d141fd9308bf8ec53e776c56e502e6c9a9aa00affdb3898

Observation 978d6b6d-be14-4b92-90fc-bc2cd0f4a49c · outbound

This paper cites Vision-language navigation with self-supervised auxiliary reasoning tasks,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Vision-language navigation with self-supervised auxiliary reasoning tasks,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T19:39:26.150581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T19:39:25.783115Z digest=sha256:a5ac9f694c79d9954120f9274f7c6f82bb5314d6cbc88648d6e558ddcdc8fc5d

Observation bc680813-c950-4f1a-a022-26eaa6d766a0 · outbound

This paper cites Bridging zero-shot object navigation and foundation models through pixel-guided navigation skill,.

Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation Bridging zero-shot object navigation and foundation models through pixel-guided navigation skill,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T19:39:25.788223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:39:25.788223Z digest=sha256:1decc60b79c69d416cd9387944eb9db32da087722b6fc8b732fbd495373b0b11

Pith citing papers

Observation ec421651-a4c8-4ccd-a647-0e7f2e8389c6 · inbound

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck cites this paper.

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:01:28.170857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-12T01:24:59.721212Z digest=sha256:b35e478734bdde6e8aaae7d424c1ccdcfb31c31b6c641fdedee9454fc2ee2709

Observation d5cf7816-9bb9-4b29-a904-b3ac9a78403c · inbound

OLIVIA: Online Learning via Inference-time Action Adaptation for Decision Making in LLM ReAct Agents cites this paper.

OLIVIA: Online Learning via Inference-time Action Adaptation for Decision Making in LLM ReAct Agents Weakly-supervised VLM-guided Partial Contrastive Learning for Visual Language Navigation

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:27:07.272715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-13T02:24:42.530103Z digest=sha256:f0bdec3e33c1cac251c6dd100f3471941b626b34afaa30d53f3dfc7a4d73c980