Pith. sign in

Paper Citation Record · LEDGER

Evaluating Vision-Language Models as Evaluators in Path Planning

As of 22 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 3 inbound Pith citation observations for arXiv:2411.18711.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.18711 v4

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:04:44.887611Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:49:07.509682Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T22:49:11.112158Z

Reference resolution

100 of 100 outbound references displayed

  • verified exact1
  • verified fuzzy59
  • unresolved39
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ade9c557-2325-44ab-8d0f-3a47a0aadaec · outbound

This paper cites Faithfulness vs.

Evaluating Vision-Language Models as Evaluators in Path Planning Faithfulness vs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.468322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.468322Z digest=sha256:89bbe94658c677944ab3818d207f1749141f201b34e1b1e2c011525bbc7d5bd8

Observation 01e605a1-7603-43b4-a035-4aa5f33450ed · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

Evaluating Vision-Language Models as Evaluators in Path Planning Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.473526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.473526Z digest=sha256:08255729498170c17eca5b5ed27f33c4b3974578747ba5981fba6efd106e2939

Observation 3761979d-6e9a-4f2d-a81c-6dd3208b7f74 · outbound

This paper cites Look Further Ahead: Testing the Limits of GPT-4 in Path Planning.

Evaluating Vision-Language Models as Evaluators in Path Planning Look Further Ahead: Testing the Limits of GPT-4 in Path Planning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.477902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.477902Z digest=sha256:1485517ced74456837bee4e871dc036dd145234748ed0bc64d93ebbe8af91bf1

Observation 8092616c-44de-4283-abe3-3c698262016e · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

Evaluating Vision-Language Models as Evaluators in Path Planning Lawrence Zitnick, and Devi Parikh

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.482092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.482092Z digest=sha256:a958fad6b635f8d588b3e3277eef02f1b48fce70c71da528266a14ff873bf30e

Observation 233da0ae-c88f-4c80-8869-17a93b546bb7 · outbound

This paper cites Vision- Language Models as a Source of Rewards, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Vision- Language Models as a Source of Rewards, 2024

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.486867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.486867Z digest=sha256:7d45b17d69e76a54469f262f33cb46f6545d5bf13ccc27262fbe97805f42edba

Observation 0c1a60d7-f548-4277-a792-8c5e1d28bd18 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.491249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.491249Z digest=sha256:c53323676d09c7e4befd430ed088c675739e3ffc4dda9baeec9a0b8ae7976bc6

Observation f4f330c3-da50-4450-ae7f-513c72441a5d · outbound

This paper cites Rehg, and Chao Zheng.

Evaluating Vision-Language Models as Evaluators in Path Planning Rehg, and Chao Zheng

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.496325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.496325Z digest=sha256:efecfbbf230652c22ff2a4e0a38a1a59eb706de662fd6ba3e5dc926967627210

Observation 32f0e98a-2ad7-40d1-932b-2028269b1a11 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

Evaluating Vision-Language Models as Evaluators in Path Planning Emerg- ing properties in self-supervised vision transformers

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.500853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.500853Z digest=sha256:40becf74521ea3a31b78fb2a5ba380b62e275fc3c2bb265048fbd5ed2425f0fc

Observation 7b5b0e98-dc65-4bf4-a698-34cf5f8e319d · outbound

This paper cites MapGPT: Map- guided prompting with adaptive path planning for vision- and-language navigation.

Evaluating Vision-Language Models as Evaluators in Path Planning MapGPT: Map- guided prompting with adaptive path planning for vision- and-language navigation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.505242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.505242Z digest=sha256:2ca78113f7e4e774dcfd12157dabdbaea7403d4df12141a7de67ad17e39a599c

Observation ca82eb8c-1164-4336-9a24-8d3c71d6786f · outbound

This paper cites Multi-object hallucination in vision language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Multi-object hallucination in vision language models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.509313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.509313Z digest=sha256:e335f094829cc5d30c5e522210c103a64f1f61d8300b4b0a8651d708cc1b3d86

Observation f65a7ef2-b024-492b-afa2-d58e5cfa6dc7 · outbound

This paper cites Autotamp: Autoregressive task and motion planning with llms as translators and check- ers.

Evaluating Vision-Language Models as Evaluators in Path Planning Autotamp: Autoregressive task and motion planning with llms as translators and check- ers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.513378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.513378Z digest=sha256:a5f9159052c29957d3ab218f4f153462bc8af39da740f4af3087c09a7ee0c69c

Observation 5b82d06f-ced8-47e9-bdd9-159ceb70355f · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

Evaluating Vision-Language Models as Evaluators in Path Planning InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.517807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.517807Z digest=sha256:52ef8f5fe3798f107e6c069747a4706921711b1d0c063e2366119ca85db34fad

Observation d99cc4f6-441a-4bd6-8385-0c5c2d47e72c · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

Evaluating Vision-Language Models as Evaluators in Path Planning Gonzalez, Ion Stoica, and Eric P

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.522611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.522611Z digest=sha256:aee767324e8dcd421a2f6fc21a8a8600316402b7af593602ff1385e732407318

Observation 587a1600-5f7f-49e0-b3d3-e84220ad8e7d · outbound

This paper cites Task and motion planning with large language models for ob- ject rearrangement.

Evaluating Vision-Language Models as Evaluators in Path Planning Task and motion planning with large language models for ob- ject rearrangement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.526554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.526554Z digest=sha256:9c64c7394f4861fc88144611f6f1c6b0c8ea60170ff0b9f5dcf1bae5d8c9bd0d

Observation 181dd57e-4680-4619-b038-02ec2e330096 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.530668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.530668Z digest=sha256:4983c08909f5b624095f56b6e98ac7d353502859a7c9c075d7f7daae18cb802b

Observation 453e7994-d40c-41bf-a3f0-180d6044a2f3 · outbound

This paper cites Tenenbaum, Leslie Pack Kaelbling, Andy Zeng, and Jonathan Tompson.

Evaluating Vision-Language Models as Evaluators in Path Planning Tenenbaum, Leslie Pack Kaelbling, Andy Zeng, and Jonathan Tompson

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.534874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.534874Z digest=sha256:a5d674e673f85442b66c3bb7d5ce9e453867ac2afb868560c647a1c5b5b405d1

Observation 58d9f6c3-7cd2-4d09-b1e9-c2e0f2cfcac8 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.538679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.538679Z digest=sha256:f80de624f72c7e6a45e715146bbb78f3c4e80c12a61724bcce1a95a5d2506bfe

Observation 36974c1e-8970-4e9f-89ad-4bac5a9e97a2 · outbound

This paper cites Cric: A vqa dataset for compositional reasoning on vision 9 and commonsense.

Evaluating Vision-Language Models as Evaluators in Path Planning Cric: A vqa dataset for compositional reasoning on vision 9 and commonsense

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.542945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.542945Z digest=sha256:1955e6672889759151c73db3bc70653e26fc47825a31f3e5f9ad9f5d22a88142

Observation fc792dc8-6b55-4f65-89b0-bfdb299ccfe9 · outbound

This paper cites Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing.

Evaluating Vision-Language Models as Evaluators in Path Planning Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.547164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.547164Z digest=sha256:b529f6f56bd4ee5c0484e43528188397b31c4cf325f4fc6faf729d15efed535a

Observation a3cf708c-c8df-4f20-ab5b-fc86919d52ee · outbound

This paper cites Vision-and-language navigation: A survey of tasks, methods, and future directions.

Evaluating Vision-Language Models as Evaluators in Path Planning Vision-and-language navigation: A survey of tasks, methods, and future directions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.551234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.551234Z digest=sha256:2a584f34042ce706bd0d88f0a75af58e522a475fe08184c1fe7fa36eacc7b180

Observation 9d340550-7e46-4412-ab55-45dd5cddecc9 · outbound

This paper cites Task success is not enough: Investigating the use of video-language models as behavior critics for catching undesirable agent behaviors.

Evaluating Vision-Language Models as Evaluators in Path Planning Task success is not enough: Investigating the use of video-language models as behavior critics for catching undesirable agent behaviors

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.554854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.554854Z digest=sha256:5ae7e99190b7dea1976d3124caef4d1faf4c17805b3e6384abd7d48e9b75b745

Observation 6fd427a0-91c6-442d-804a-4ff8ed4cb593 · outbound

This paper cites Hal- lusionbench: An advanced diagnostic suite for entangled language hallucination & visual illusion in large vision- language models, 2023.

Evaluating Vision-Language Models as Evaluators in Path Planning Hal- lusionbench: An advanced diagnostic suite for entangled language hallucination & visual illusion in large vision- language models, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.178934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.558971Z digest=sha256:8630eb49fb4562e303fa3578f2111a566d1ba2b1518b1c8d4cef6d4e78f1508a

Observation 16a0fdbb-a7fb-440a-b308-e0cd982e0b36 · outbound

This paper cites Generating and evolving reward functions for highway driving with large language models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Generating and evolving reward functions for highway driving with large language models, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.164228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.563253Z digest=sha256:0ce88af92f149c3c8e2f30b857cfd862903e331151512fdca000452f930c4eb8

Observation da16eae5-a81d-4993-9316-ebe64dc9fb09 · outbound

This paper cites Hudson and Christopher D.

Evaluating Vision-Language Models as Evaluators in Path Planning Hudson and Christopher D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.149923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.567498Z digest=sha256:ffd02d450ac0aa63635a46912c615694f78ff5b7fe4b356a6971d4852aa371cd

Observation 751f3513-7d07-40c2-a2ec-6a1309b9c47e · outbound

This paper cites Open- clip, 2021.

Evaluating Vision-Language Models as Evaluators in Path Planning Open- clip, 2021

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.134842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.571709Z digest=sha256:aaebb052f4ff97d787ff58047e501ea97f53bf23e5febc4fe30a6750d2b31be2

Observation 104ad04a-0b97-4ab9-97f6-ae2a612b739d · outbound

This paper cites What’s ”up” with vision-language models? Investigating their strug- gle with spatial reasoning.

Evaluating Vision-Language Models as Evaluators in Path Planning What’s ”up” with vision-language models? Investigating their strug- gle with spatial reasoning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.121589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.575741Z digest=sha256:0f5b6bdbe4d5b56ec5b56f1d19a83581c42cc24867fc499d431a6d307653d938

Observation 7f3b778d-2e92-45b4-b4a3-c618f13bdf27 · outbound

This paper cites Position: LLMs Can’t Plan, But Can Help Planning in LLM-Modulo Frameworks.

Evaluating Vision-Language Models as Evaluators in Path Planning Position: LLMs Can’t Plan, But Can Help Planning in LLM-Modulo Frameworks

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.107670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.579754Z digest=sha256:aecc03c3c961924d6e86db71b7db602069a63c5176de20b09866693b5a63684e

Observation f746ba24-f58f-4904-a87b-75630d9f239d · outbound

This paper cites Kavraki, P.

Evaluating Vision-Language Models as Evaluators in Path Planning Kavraki, P

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.093809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.584135Z digest=sha256:d72edc4b07baa956d4e017b85a3817e47ef9482604c626ad08ca3033a894467f

Observation c5920d40-9111-482c-8d24-0b7ee755258c · outbound

This paper cites Kuffner and S.M.

Evaluating Vision-Language Models as Evaluators in Path Planning Kuffner and S.M

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.080404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.588280Z digest=sha256:5b24a5490ec46514b52c3fa8d0d1cf23b03fc938d1d13167a88776f72880fa8d

Observation 8a08d27a-0f30-4d6b-89e9-d0d04968ecaa · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.592439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.592439Z digest=sha256:12f1538f958e32eff44969c9cd986859648e6041e83d9d98e5936d98ccb11483

Observation 72860de6-2c21-4c95-bc23-a7976209dd30 · outbound

This paper cites Why i’m optimistic about our align- ment approach: Evaluation is easier than generation.

Evaluating Vision-Language Models as Evaluators in Path Planning Why i’m optimistic about our align- ment approach: Evaluation is easier than generation

Reference 31

Resolution
verified exact
raw_fallback, observed 2026-08-12T11:04:45.048905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.596527Z digest=sha256:ade39be38621358dab76b7b1534766d1b81b34eaecbbe2a3f325701ca84434da

Observation d63578c3-077e-4918-a307-b20b62015e44 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Evaluating Vision-Language Models as Evaluators in Path Planning LLaVA-OneVision: Easy Visual Task Transfer

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.605211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.605211Z digest=sha256:fe70936dc776ec8094ee13acb7cded7918d2ab7d4e509b158e8840cc4940138e

Observation ed81fadc-07f5-49df-801b-2f5cd7737962 · outbound

This paper cites Auto mc-reward: Automated dense reward design with large language models for minecraft.

Evaluating Vision-Language Models as Evaluators in Path Planning Auto mc-reward: Automated dense reward design with large language models for minecraft

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.045353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.610072Z digest=sha256:8e9557076ecf60934cdd2d46e65746243eb0c218ffe8efba618640da519dc63a

Observation 28305202-3f6b-485e-9d9c-342077825d35 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Evaluating object hallucination in large vision-language models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.614434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.614434Z digest=sha256:ca7c3b86407b9d864bd9041e3510c2b979c476ddf425a45cc4752b3e6f78e069

Observation 7c5da2e7-6986-4049-aac8-8976bd59f1b2 · outbound

This paper cites OmniBench: Towards The Future of Uni- versal Omni-Language Models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning OmniBench: Towards The Future of Uni- versal Omni-Language Models, 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.021161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.618780Z digest=sha256:8a036454d06b798a8cc1cc1fdac9f99e1585097a33360fd7156a1df073f4d7c5

Observation 69284036-e248-469c-85fc-d829989cd295 · outbound

This paper cites Chin, Shuhong Chai, Neil Bose, and Eonjoo Kim.

Evaluating Vision-Language Models as Evaluators in Path Planning Chin, Shuhong Chai, Neil Bose, and Eonjoo Kim

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:46.007403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.622940Z digest=sha256:80e5bb4c639449654c4aa1f976176fd21c4e20d31fd3e788c52e2c08c3ac7795

Observation 0c3e868d-fd36-49f4-8691-3b309339f336 · outbound

This paper cites Visual instruction tuning.

Evaluating Vision-Language Models as Evaluators in Path Planning Visual instruction tuning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.626964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.626964Z digest=sha256:3de980b652ea0ee645e9e4b250e43349047c61d634781b425cec8bffc94ab6c4

Observation 4a300cf0-fe05-4baf-8e0e-f722440854cb · outbound

This paper cites Improved Baselines with Visual Instruction Tuning, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Improved Baselines with Visual Instruction Tuning, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.984335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.631016Z digest=sha256:fefa7cb7377d61c1b13cdd80d5892fcb8fada213dc975dc77d490c05bd8161d7

Observation c5870301-fb45-40dc-bf24-19ac27975806 · outbound

This paper cites LLaV A-NeXT: Im- proved reasoning, OCR, and world knowledge, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning LLaV A-NeXT: Im- proved reasoning, OCR, and world knowledge, 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.970141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.635085Z digest=sha256:40573b06214362e899005542e41f09e4dc8185b786ce4e37d62d379f09345b7a

Observation 3534786f-a71e-428d-8782-29cc05eb01c3 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? In Computer Vi- sion – ECCV 2024 , pages 216–233, Cham, 2025.

Evaluating Vision-Language Models as Evaluators in Path Planning Mmbench: Is your multi-modal model an all-around player? In Computer Vi- sion – ECCV 2024 , pages 216–233, Cham, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.954628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.639816Z digest=sha256:d751cc62eab03e62acab57a6ab323b3b6ce2d0ae4b58a10dcc927465acfbc2a3

Observation d1788933-3460-4a41-a68d-02d1fdcca032 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.941485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.644548Z digest=sha256:bdb9bf9f04391893e1451275fe356c5755b511032b62f5918eac4d517feaa9f9

Observation 02298d95-bdd7-4554-879d-6161e2687468 · outbound

This paper cites Mathvista: Evaluating mathe- matical reasoning of foundation models in visual contexts.

Evaluating Vision-Language Models as Evaluators in Path Planning Mathvista: Evaluating mathe- matical reasoning of foundation models in visual contexts

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.928484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.649163Z digest=sha256:0157c90cd72dcc50699ead52f5aec181f066445bfb8e04fc53c74dd594e27d86

Observation 9545c317-4529-403b-a694-402e8e14dd65 · outbound

This paper cites Ok-vqa: A visual question answering benchmark requiring external knowledge.

Evaluating Vision-Language Models as Evaluators in Path Planning Ok-vqa: A visual question answering benchmark requiring external knowledge

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.914836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.653818Z digest=sha256:6e54590eae47c22292a8ebd46e3d7e4b4cbf1ac757fadb8f4968139d74d59285

Observation fc42eb96-b56d-4b35-92c0-c43a37e5218d · outbound

This paper cites LLM-a*: Large language model en- hanced incremental heuristic search on path planning.

Evaluating Vision-Language Models as Evaluators in Path Planning LLM-a*: Large language model en- hanced incremental heuristic search on path planning

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.900899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.657811Z digest=sha256:527d5831276af33a742afcc63c59eca6ff2ebb6b8b91d757d0a847e9319ac2c4

Observation f2778f7b-f303-43de-a0d0-90cfa4fef962 · outbound

This paper cites Jmmmu: A japanese massive multi- discipline multimodal understanding benchmark for culture- aware evaluation, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Jmmmu: A japanese massive multi- discipline multimodal understanding benchmark for culture- aware evaluation, 2024

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.877713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.665547Z digest=sha256:2f3134db9ce61f6a1cd2d503a543fa207ce2bbff7d7e3f1dd2a28898ab6ccb5e

Observation df49d972-79f8-4305-a284-f22d0b8bb3c8 · outbound

This paper cites GPT-4V(ision) System Card.

Evaluating Vision-Language Models as Evaluators in Path Planning GPT-4V(ision) System Card

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.864318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.669617Z digest=sha256:e2d1f048e5cf411d5f5d2874ab390fbae02a7248f524a17778f110b67922a16b

Observation 3a46277f-1688-468b-ae25-4576b0b7248c · outbound

This paper cites GPT-4o System Card, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning GPT-4o System Card, 2024

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.852234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.674070Z digest=sha256:da7da970562e136b004d61ead9a9db179569a8c1e2bdc9fc478bcdcb081ecda5

Observation 49af364b-1ec3-43b6-a1ab-1c1960625206 · outbound

This paper cites GPT-4 Technical Report, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning GPT-4 Technical Report, 2024

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.839972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.677668Z digest=sha256:d138e31cde9704f7d2a88f3f5b63bc402920563fb36a85a22878a3f1a7f354c2

Observation 2be7e609-e614-41a6-8640-3b242c9a3142 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.826918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.681380Z digest=sha256:2c14fc1c7866c96860e48dfa8b94cfa57949466ea530eef73b4aebff947b285f

Observation c0808b7a-de1c-4e46-a2b1-b2ba824cbcb5 · outbound

This paper cites LangNav: Lan- guage as a perceptual representation for navigation.

Evaluating Vision-Language Models as Evaluators in Path Planning LangNav: Lan- guage as a perceptual representation for navigation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.813938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.685785Z digest=sha256:7d8c2062b2b00028d6732e722f69d17e6afc7bb974dbeaa4447e59ad73384f2f

Observation fd4834ce-28e1-4ef1-bf36-d618db0fb3bb · outbound

This paper cites VLP: Vision Language Planning for Autonomous Driving.

Evaluating Vision-Language Models as Evaluators in Path Planning VLP: Vision Language Planning for Autonomous Driving

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.800745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.690397Z digest=sha256:0aaf1ee8c31d44e15ec4c1e29ee5ef510a5a68bdd345be6c80e859b1bfe6dbdb

Observation 99c8b953-1c46-4ad5-b9d8-5260e87b8c2e · outbound

This paper cites Path planning for autonomous underwater vehicles.

Evaluating Vision-Language Models as Evaluators in Path Planning Path planning for autonomous underwater vehicles

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.786723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.694719Z digest=sha256:7588fef25778152c9eee0132830c9fb28ea9e2c23b844362e91a3f3c91e26119

Observation c1538b26-955b-4445-88a4-12e2ce8e465c · outbound

This paper cites Clearance- driven motion planning for mobile robots with differential constraints.

Evaluating Vision-Language Models as Evaluators in Path Planning Clearance- driven motion planning for mobile robots with differential constraints

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.774346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.699192Z digest=sha256:0d3d26a1eaa2a74d8e5bde5c5c85c428534123d6f4da6f78cd045f7f58aee6e2

Observation 7005ad3b-978d-46f4-863b-f689888b76a5 · outbound

This paper cites Plonski, Pratap Tokekar, and V olkan Isler.Energy- Efficient Path Planning for Solar-Powered Mobile Robots , pages 717–731.

Evaluating Vision-Language Models as Evaluators in Path Planning Plonski, Pratap Tokekar, and V olkan Isler.Energy- Efficient Path Planning for Solar-Powered Mobile Robots , pages 717–731

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.761555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.703585Z digest=sha256:035cba6cde2e7a65225f31a49553fad9c9bdb45227f9180d94256c70030c3dff

Observation 97acd0c8-db5d-457b-835e-ca4a44c14580 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Evaluating Vision-Language Models as Evaluators in Path Planning Learning transferable visual models from natural language supervision

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.748543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.708095Z digest=sha256:be038f599b68f51c763baffa33aec6197bbc134c5812ab777e2062359ed5f386

Observation 208c4b37-928b-43b2-9bb2-3b32290f0cb2 · outbound

This paper cites LAION-5b: An open large-scale dataset for train- ing next generation image-text models.

Evaluating Vision-Language Models as Evaluators in Path Planning LAION-5b: An open large-scale dataset for train- ing next generation image-text models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.734649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.712367Z digest=sha256:c6d81e792eeed8058b6f2be088c6e4904dbac9dcf080e33467c60c19dc1e2455

Observation c8c31025-8e85-4d14-93d6-13da16e0cf44 · outbound

This paper cites Investigating the Limitation of CLIP Mod- els: The Worst-Performing Categories, 2023.

Evaluating Vision-Language Models as Evaluators in Path Planning Investigating the Limitation of CLIP Mod- els: The Worst-Performing Categories, 2023

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.719336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.716271Z digest=sha256:6833a36641a9f93146f3edca2edee60b621e2fd00d90661af1d02eb18e08c069

Observation 1e0c459e-dbdb-4c73-b732-7123298430c9 · outbound

This paper cites Tenen- baum, Leslie Pack Kaelbling, and Michael Katz.

Evaluating Vision-Language Models as Evaluators in Path Planning Tenen- baum, Leslie Pack Kaelbling, and Michael Katz

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.705388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.720569Z digest=sha256:f9f9bd9cbd86999fc668d7139ae314f4572d7d42853529b1434fd8c6105fbb88

Observation 13b5b4a4-0daa-4350-ad8e-11a710779da2 · outbound

This paper cites Sucan, Mark Moll, and Lydia E.

Evaluating Vision-Language Models as Evaluators in Path Planning Sucan, Mark Moll, and Lydia E

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.579131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.724836Z digest=sha256:fb6b0660c448ed3a44ab9a039958834350a823fd0edb413a506f8f5277231f07

Observation 51a45461-d8de-48fc-8e17-dbde915b2d2c · outbound

This paper cites Explore the hallucination on low-level perception for mllms, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Explore the hallucination on low-level perception for mllms, 2024

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.565304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.729106Z digest=sha256:121579559ef8e685b14fbcec9cf16868b70ee9290cdc398984f509e2843ea352

Observation 0c054946-0e83-4f26-b2a6-214da8be18a7 · outbound

This paper cites Learning to nav- igate unseen environments: Back translation with environ- mental dropout.

Evaluating Vision-Language Models as Evaluators in Path Planning Learning to nav- igate unseen environments: Back translation with environ- mental dropout

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.550962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.734506Z digest=sha256:2dbfb8e9d91227f49773eb5dc47da36d3969bbb9f3edbdf4d7b4a0bb443df347

Observation 338e3d43-c4cd-4d4d-814a-ce43c9e89ffe · outbound

This paper cites Gemini: A family of highly capable multi- modal models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Gemini: A family of highly capable multi- modal models, 2024

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.738697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.738697Z digest=sha256:56fe35d53e2fc20800cf2703e6af2aad894c31825a47090851853dbf09744959

Observation 8d596672-66fd-4258-8223-56030596b5dd · outbound

This paper cites Mass- producing failures of multimodal systems with language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Mass- producing failures of multimodal systems with language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.526822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.743101Z digest=sha256:1d15800f91627f15fc0db1ecf74322485581091cf14d3dd3d9688ce87c8b9fac

Observation 38fbf9ff-1a3a-4dee-87a6-5bdf75093e3f · outbound

This paper cites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs, 2024

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.513386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.747472Z digest=sha256:098dd7973d2e96ff35c6f11e2d2e1af4f4e49ecb72d8452485cc745f922f8ce4

Observation 2e9a1d4a-dc7f-44f4-8186-41d648aa540e · outbound

This paper cites Llama 2: Open foundation and fine- tuned chat models, 2023.

Evaluating Vision-Language Models as Evaluators in Path Planning Llama 2: Open foundation and fine- tuned chat models, 2023

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.500389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.751680Z digest=sha256:0a6c57999f728fea3406dc741fd3adee99f98ffd438373347b6a61560f6595f2

Observation e3c0274e-1378-4ba9-8692-4305b46a184d · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.487491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.756200Z digest=sha256:03acc445d216bfe2c928da325de2a4ef93ddf64a6d0b99236de9c1c95089bac1

Observation 8218cfe2-8240-4893-beb4-ca100ef9707a · outbound

This paper cites Large Language Models Still Can’t Plan (A Benchmark for LLMs on Planning and Reasoning about Change).

Evaluating Vision-Language Models as Evaluators in Path Planning Large Language Models Still Can’t Plan (A Benchmark for LLMs on Planning and Reasoning about Change)

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.474491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.760423Z digest=sha256:e440245b0f618ac3090483420eb97123249edada96facc60299fd7b5e0d6730e

Observation a5caeb81-5348-4536-9ab7-fec37bbde83b · outbound

This paper cites LLMs Still Can’t Plan; Can LRMs? A Preliminary Evaluation of OpenAI’s o1 on PlanBench, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning LLMs Still Can’t Plan; Can LRMs? A Preliminary Evaluation of OpenAI’s o1 on PlanBench, 2024

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.461224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.764554Z digest=sha256:3bbe1e8ef62dae6a73d949d3f177646a746f6c23aaa3c9a0841372dfbdc2c268

Observation a70f9c6a-7463-475c-9399-92e807d2b61c · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Evaluating Vision-Language Models as Evaluators in Path Planning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.768564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.768564Z digest=sha256:75d616b63d1a9aaf01e5043e9690657dc135c8c97d670638da9966f7fcf09b58

Observation e6ef1e45-00f8-4e60-bb4d-7d12c52c19e6 · outbound

This paper cites Reinforced cross-modal matching and self- supervised imitation learning for vision-language navigation.

Evaluating Vision-Language Models as Evaluators in Path Planning Reinforced cross-modal matching and self- supervised imitation learning for vision-language navigation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.446679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.773169Z digest=sha256:196d7d2b12e838950a4faea083ef948bce97836e2342991c36479d818505dfc0

Observation a92fed30-c6c7-4975-9a0d-0e9b45721379 · outbound

This paper cites Chi, Quoc V Le, and Denny Zhou.

Evaluating Vision-Language Models as Evaluators in Path Planning Chi, Quoc V Le, and Denny Zhou

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.430518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.777491Z digest=sha256:c0eb849432a71f40bc1de9139054e41906b3f6105561d17f2f48a74a68602b95

Observation 61302c78-1f1e-406c-8c9b-caa0d484fed0 · outbound

This paper cites Clip-dinoiser: Teaching clip a few dino tricks for open- vocabulary semantic segmentation.

Evaluating Vision-Language Models as Evaluators in Path Planning Clip-dinoiser: Teaching clip a few dino tricks for open- vocabulary semantic segmentation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.416581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.781477Z digest=sha256:4cac187490111dbd8098455cb3f8b8847ed118d3b83b735cbecea92611be4add

Observation 6545e40f-d0ed-4e12-9072-8eb767b47fe9 · outbound

This paper cites Text2Reward: Reward Shaping with Language Models for Reinforcement Learning.

Evaluating Vision-Language Models as Evaluators in Path Planning Text2Reward: Reward Shaping with Language Models for Reinforcement Learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.402513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.785562Z digest=sha256:69d50d815ad269288659b2e33efb152f703448d068d81b2bb9c34fa1c9220cfb

Observation 34eddb6d-510e-41dd-9b97-90abb6b76ff3 · outbound

This paper cites Evaluating spatial understanding of large language models.

Evaluating Vision-Language Models as Evaluators in Path Planning Evaluating spatial understanding of large language models

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.387997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.789884Z digest=sha256:897f65c5b4cb6485e17bcbcf2f28e4c1f0183f81cffa85ff1fdf04d905e56ecf

Observation bd391b2f-403e-4ef6-8dd1-074afec3a009 · outbound

This paper cites Guiding long-horizon task and motion planning with vision language models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Guiding long-horizon task and motion planning with vision language models, 2024

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.373817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.793901Z digest=sha256:c133e5bac2d18a75beda03e5c89098d0a1e1904360abbe6e938bd24e907a1035

Observation 606d6a23-387b-4cc5-b586-367b13321856 · outbound

This paper cites Coca: Contrastive captioners are image-text foundation models.

Evaluating Vision-Language Models as Evaluators in Path Planning Coca: Contrastive captioners are image-text foundation models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.797911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.797911Z digest=sha256:274bf83740e251b57e65e9e492bb45a6c2ea2f29f01631f45b0328c7eef72abd

Observation b60f5fd1-affd-40a6-9a94-6ddabd629ecd · outbound

This paper cites Mm-vet: Evaluating large multimodal models for integrated capabilities.

Evaluating Vision-Language Models as Evaluators in Path Planning Mm-vet: Evaluating large multimodal models for integrated capabilities

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.349494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.801895Z digest=sha256:df7a6fb203cbbe415af79d58955cc5c386e72ba951f7b436f2bd6ad54c936031

Observation 1faf5b72-fe6e-4f91-8347-7cf89c35e0b0 · outbound

This paper cites MMMU: A Massive Multi-discipline Multimodal Under- standing and Reasoning Benchmark for Expert AGI.

Evaluating Vision-Language Models as Evaluators in Path Planning MMMU: A Massive Multi-discipline Multimodal Under- standing and Reasoning Benchmark for Expert AGI

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.335634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.805741Z digest=sha256:5636d1986c69f475e2ec5a011d2bdfae7b0aed7536effbd4c71243e3be1c038a

Observation 913e0fea-1979-4c77-8c4f-94073a936305 · outbound

This paper cites MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark.

Evaluating Vision-Language Models as Evaluators in Path Planning MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.809641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.809641Z digest=sha256:f2777f12a8c0a2a26c52bd4a631912570a9eff8b5e76e1832af9e7b6558a9408

Observation 42b58dfc-dacc-4079-8477-7d891a1bbed2 · outbound

This paper cites Sigmoid loss for language image pre-training.

Evaluating Vision-Language Models as Evaluators in Path Planning Sigmoid loss for language image pre-training

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.321121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.813434Z digest=sha256:491203aa615422c41c69f6b9fbccd648bb12c205185084ecbea12b67c19ac3e9

Observation 92df6538-c39a-4af8-b197-da54b87e7d31 · outbound

This paper cites CMMMU: A Chinese Massive Multi-discipline Multi- modal Understanding Benchmark, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning CMMMU: A Chinese Massive Multi-discipline Multi- modal Understanding Benchmark, 2024

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.307576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.817313Z digest=sha256:73d9a29d4456988468b7143af4e036b7505dc3e1bcb6dec1c459f0035a96c675

Observation ee29249f-ac69-4f5b-9198-d65cce5efafb · outbound

This paper cites Multimodal chain-of-thought rea- soning in language models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Multimodal chain-of-thought rea- soning in language models, 2024

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.295221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.821187Z digest=sha256:2190ec327743699585a06f5fc50eb0385265614229bbdd1c17cff49608a0d169

Observation b9ed1ede-e31e-49eb-95b6-cc626a4cc5cf · outbound

This paper cites Policy Improvement using Language Feed- back Models, 2024.

Evaluating Vision-Language Models as Evaluators in Path Planning Policy Improvement using Language Feed- back Models, 2024

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.282719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.824729Z digest=sha256:16efcffabd9f6b04ed50d8e96d12a96fafce22937bd849fce0e4e2c0162fb1c8

Observation 3898e54f-21ac-40e0-97ea-31d3d2feb1a5 · outbound

This paper cites Visual7W: Grounded Question Answering in Images.

Evaluating Vision-Language Models as Evaluators in Path Planning Visual7W: Grounded Question Answering in Images

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.269706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.828765Z digest=sha256:02ab9e0bddaea0e9a0af6924f394c96d51227dcb93b752b618d0c2130d3cc692

Observation 30afc4de-d523-4a54-938e-65a6753c89cb · outbound

This paper cites Min Clearance = min pj ∈P min Oi∈O D(pj, Oi).

Evaluating Vision-Language Models as Evaluators in Path Planning Min Clearance = min pj ∈P min Oi∈O D(pj, Oi)

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.256300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.832912Z digest=sha256:dd390ce7c2ba8ad1e5b9b5ab66107c46e615cf69958cf2bb44228349b4651da5

Observation bba82f7d-5a91-44ef-a3e9-71d7261c0db1 · outbound

This paper cites Max Clearance = max pj ∈P min Oi∈O D(pj, Oi).

Evaluating Vision-Language Models as Evaluators in Path Planning Max Clearance = max pj ∈P min Oi∈O D(pj, Oi)

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.243179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.837254Z digest=sha256:bf413a7396481c632afc324e7b1225c1054f69a3abeba4326a7b0502a06bb4d6

Observation 7ba76492-d207-4345-a47e-c5540837a059 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.229295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.841771Z digest=sha256:81e20337e1a8e03581614e0e2fc10899c80b439c2ec18d0870b619da03c81759

Observation 71a11df0-45e8-480c-b18d-79f385857564 · outbound

This paper cites Path Length = nX j=2 D(pj−1, pj).

Evaluating Vision-Language Models as Evaluators in Path Planning Path Length = nX j=2 D(pj−1, pj)

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.215081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.845673Z digest=sha256:2a9beca3fc14ec99ccf4aee3c4f0b240ab74a3d53436b4926c2a653606b84212

Observation 1482156a-5a18-4662-bd23-dfddcca7b0f6 · outbound

This paper cites Smoothness = nX j=3 θj where θj is the angle between the vectors − − − − − − →pj−2pj−1 and− − − − →pj−1pj.

Evaluating Vision-Language Models as Evaluators in Path Planning Smoothness = nX j=3 θj where θj is the angle between the vectors − − − − − − →pj−2pj−1 and− − − − →pj−1pj

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.201026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.849758Z digest=sha256:770bab59e027cf4130f39b6e662237433b0744bbee575d0f3a4a23f0e80c7c1e

Observation 585fbad9-0e2f-4a7f-a1d7-0bafe47bd97c · outbound

This paper cites Sharp Turns = nX j=3 δj, where δj = ( 1 if θj > 90◦ 0 otherwise.

Evaluating Vision-Language Models as Evaluators in Path Planning Sharp Turns = nX j=3 δj, where δj = ( 1 if θj > 90◦ 0 otherwise

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.187140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.854193Z digest=sha256:8f8e4cb3febf034f232aa56ac6fcda3a9d9d80dce066ea5dc5f290a64272d11c

Observation a07297e6-002d-4573-b2b1-06743bbd9168 · outbound

This paper cites Maximum angle = n max j=3 θj C.

Evaluating Vision-Language Models as Evaluators in Path Planning Maximum angle = n max j=3 θj C

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.173624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.858557Z digest=sha256:2ef4ec55a01855479439ff87b5e18a60c4edcfdb92c9f8349449cac75f874e1c

Observation f1989607-864e-4dc0-b6d3-9f660d93f832 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.160650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.863056Z digest=sha256:2855a0a46a7e1c8e9e3fff15a196e0d37a3657b324c0ac6ca0ef50b3dd2eca91

Observation 8d3f074f-d807-4773-a14e-7c25a4875a3e · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.146507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.867141Z digest=sha256:df512d55a85aa52211b0afba16c4348bd6914da0772e45e44ef18bc10f24bad5

Observation 1a7f99d8-acec-4c38-ae02-b07a8bf41e79 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.130946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.871269Z digest=sha256:e3bfa2621f021851598a2c6a281605b4894b9c5e70af87f2ce39ee096503668e

Observation 9c553200-9c59-4854-a2ec-4cd40bffa397 · outbound

This paper cites Smoother paths have a lower smoothness value.

Evaluating Vision-Language Models as Evaluators in Path Planning Smoother paths have a lower smoothness value

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.117915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.875172Z digest=sha256:4d0dd6c4162f5036868cba2efaee0736acdaa91838ba586303bb39977179bd15

Observation 90985cf5-5f42-43b0-988f-9f1e9e4d1cb7 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 98

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.104257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.879270Z digest=sha256:49d49457f5dcd87cf20da72e7714c5e14fbe4ade0efda1d762e64c5a8dfa55fc

Observation 252f6b04-5df5-4c9a-8607-d1f62ad1f11d · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:04:45.091181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.883284Z digest=sha256:54b9d9087541f33e7e51841238ec2e9c50e45e2383e5bdc51285029bf0c91131

Observation d23d5a32-8b22-4c8a-bb54-0cfab950f2be · outbound

This paper cites Path 1 has a smaller value for the given met- ric.

Evaluating Vision-Language Models as Evaluators in Path Planning Path 1 has a smaller value for the given met- ric

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:04:45.077702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.887611Z digest=sha256:5d481ae4e6498db27e22adb53807790a7bc3c32fbb0eede792331541c5b17552

Observation a058d5c1-a1b5-4368-b40c-1362551359d5 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 2022

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T11:04:46.058456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T11:04:44.600983Z digest=sha256:19f84a7344535e6db72079d12885c54e3218733c1c3d9d4ae1acc8edac0fe4d2

Observation da70871a-eec8-4526-9040-f859f5465cf5 · outbound

This paper cites an unresolved cited work.

Evaluating Vision-Language Models as Evaluators in Path Planning Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-12T11:04:44.661729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:04:44.661729Z digest=sha256:4c294eb0a38fc52dd9567105133871c035d424889cb49d70e937e8d179f59bac

Pith citing papers

Observation 1a90adff-d994-4835-914a-c1ef00a2389d · inbound

HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction cites this paper.

HRIBench: Benchmarking Vision-Language Models for Real-Time Human Perception in Human-Robot Interaction Evaluating Vision-Language Models as Evaluators in Path Planning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:49:11.200446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T22:49:07.509682Z digest=sha256:723734df934574f64bc0b9449d05a09607c0052603657fb1e1bfc0521d650bf1

Observation 78728900-4e2e-484e-b0cf-c68cb9e86c6b · inbound

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach cites this paper.

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach Evaluating Vision-Language Models as Evaluators in Path Planning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T16:52:49.914971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:52:49.914971Z digest=sha256:f25a7cef24eb1723e419963864d9fda25d472896550f584bc4e4f5c9bfcddb0e

Observation 26b5b30f-c90a-4d2c-aaad-dd4e026f029c · inbound

Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies cites this paper.

Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies Evaluating Vision-Language Models as Evaluators in Path Planning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T07:49:44.998466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:49:44.998466Z digest=sha256:57fb9ce2d97fc3b1d6c4928cdddd6afa2627a1fb79d35217b02d9486cd0c31b6