Pith. sign in

Paper Citation Record · LEDGER

Visual Grounding in Zero-Shot Vision-Language Control

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2608.06154.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06154 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:12.082156Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a603f8c3-022b-4147-8169-63fe57346953 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Visual Grounding in Zero-Shot Vision-Language Control Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.790543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:08.270889Z digest=sha256:86411b5934342c1a6d2ec04b22720c1aafde29bb0079a43dd3a54800a82f1af6

Observation 0e43c405-639a-4c5b-b059-ff170912dad6 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Visual Grounding in Zero-Shot Vision-Language Control Improved Baselines with Visual Instruction Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.301524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.301524Z digest=sha256:84e932ee38a77e8cb344f2894a52d40d28c0934fca7b647c59af4f827efed743

Observation 4603868a-d4e8-4c73-89f7-1d53b5515177 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.373123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.373123Z digest=sha256:3807b07c68ca1af041c84c53bb93720c484e078403b86396e09fa0b01e864319

Observation a1a210d9-32bb-4b99-8083-5b762cb6b2ef · outbound

This paper cites Qwen2.5-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.442796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.442796Z digest=sha256:405cd47fe1312703ca5a0ee0839357bd6ac9712f1c01339398b60fcc0532e93b

Observation dbb8b77a-3727-4987-a03a-e8a06282aa5f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.553499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.553499Z digest=sha256:82c9d9b12a537c1696d56addb1e9effbeb9f44f3ffb2ea8a9c89eca9639dd434

Observation fb1a59d1-ed3d-4bad-9ccb-6af984254bd5 · outbound

This paper cites Qwen3-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.643309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.643309Z digest=sha256:603643603d3c7cf6648967a6ed8ca925eacf99f1648600c887571ecc54afb9a5

Observation 4d55ec77-a452-4a5a-becb-390200596896 · outbound

This paper cites Gemma 4 Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Gemma 4 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.729258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.729258Z digest=sha256:9352dacd1f5757c4241297a5c3e25e2a49168bb6414e7ca5fcce3052aa6e0ce4

Observation ec3159d7-5976-4240-90d2-7389098270cd · outbound

This paper cites Qwen3.5: Towards native multimodal agents,.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3.5: Towards native multimodal agents,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.579700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:08.814585Z digest=sha256:f1c49ee564277e68489dde4a202fa9fd8c587edfd8a4fef66b41f6f89cc13c4f

Observation f585939b-9dd9-4c7b-88d2-52e812ad10b1 · outbound

This paper cites Introducing Mistral 3,.

Visual Grounding in Zero-Shot Vision-Language Control Introducing Mistral 3,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.388390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:08.880443Z digest=sha256:c94f593b36efe72ee3c3991e3884e4b200115ab1b0e08df05f03147a1bd30e77

Observation 06420827-4300-4bed-81da-c6f12310a09a · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.961126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.961126Z digest=sha256:21c8ec69f824394c9384975e3798e9f829f26f2a60a001646a9c93cd7ca9ae9c

Observation 9ffb7a4b-5bc3-433f-b84e-657efa32534c · outbound

This paper cites SmolVLM: Redefining small and efficient multimodal models.

Visual Grounding in Zero-Shot Vision-Language Control SmolVLM: Redefining small and efficient multimodal models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.055707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.055707Z digest=sha256:2ec0be2649e96bf2f5e15cb3269bbd48a73f1d25bf98cf223fbca97c671d581d

Observation 2e96abb8-7e75-49cf-bad1-44b987d96f15 · outbound

This paper cites NavGPT: Explicit reasoning in vision- and-language navigation with large language models,.

Visual Grounding in Zero-Shot Vision-Language Control NavGPT: Explicit reasoning in vision- and-language navigation with large language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.185974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.150370Z digest=sha256:e51563d747d002e059d006057364eeb7c03463149c6a5c12dff8640dbafc703a

Observation 2ac4a8d0-ddbe-4845-83ab-753e13619f30 · outbound

This paper cites VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.988861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.191891Z digest=sha256:3969212f30abf84584a21699e992060e7e14a915dd93c8bf462034d263f650b7

Observation 1617f2b0-e8b3-4d04-a32d-24e9620e142a · outbound

This paper cites VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.794638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.271864Z digest=sha256:31d61f5a5e5334f6eefd3314b9295cf224d0c34ffb82ea0215665b7666001c00

Observation f31e68ee-2539-4c67-bb19-0ef0a3340012 · outbound

This paper cites MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,.

Visual Grounding in Zero-Shot Vision-Language Control MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.624658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.376035Z digest=sha256:1cf797ca965fbbf08fd9d4b6bb940347f2dc1f1da70930e90f55b6685e5c06fd

Observation 70fc3451-2ab0-43bf-bfdb-0fc46a2249cf · outbound

This paper cites HazardVLM: A video language model for real-time hazard description in automated driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control HazardVLM: A video language model for real-time hazard description in automated driving systems,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.405609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.469350Z digest=sha256:7ec9ad37916a2ef9aa873052e19ac6991167f7dce63a4c98f3cf03809db36e97

Observation 25951574-ea08-42cf-87f1-209bb847b839 · outbound

This paper cites LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,.

Visual Grounding in Zero-Shot Vision-Language Control LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.178944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.543585Z digest=sha256:ed720a88a2d6e3be51fc3f039e0ba57dca286434964792d086f539ba051c2aa1

Observation c5278006-8a5e-4c69-99f3-2bcd668bf024 · outbound

This paper cites Semantic scene understand- ing with large language models on unmanned aerial vehicles,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic scene understand- ing with large language models on unmanned aerial vehicles,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.962984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.642706Z digest=sha256:ee96c57ce1989c5d5ed27478f4f12c5787279c2576740ca3c24ae1b387aa2a27

Observation 91026761-6ef7-4e28-b7b5-ce3798e118e0 · outbound

This paper cites Shortcut learning in deep neural networks,.

Visual Grounding in Zero-Shot Vision-Language Control Shortcut learning in deep neural networks,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.715223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.715223Z digest=sha256:6ff79286fbc846df871d8908da57c2a1110ead15247963249512ca3d8e349be3

Observation f29b433f-6864-482b-8456-418ec2180865 · outbound

This paper cites Beyond accuracy: Behavioral testing of NLP models with CheckList,.

Visual Grounding in Zero-Shot Vision-Language Control Beyond accuracy: Behavioral testing of NLP models with CheckList,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.751966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.835083Z digest=sha256:c1697b1e6a3584853d82011b47e05413ee8e9ab30e8b487b583cd1c04fcb5415

Observation ca7c9e58-04a9-4a16-9f56-56cadde12161 · outbound

This paper cites Holistic Evaluation of Language Models.

Visual Grounding in Zero-Shot Vision-Language Control Holistic Evaluation of Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.972585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.972585Z digest=sha256:dda1155747e17f2e8a1170ba34b772eb61e265a6a4a31d6dbf4b03e295212312

Observation e00c28b0-ce6a-4c24-a854-1d9bf318dca4 · outbound

This paper cites Metamorphic testing: A review of challenges and opportunities,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing: A review of challenges and opportunities,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:10.087742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:10.087742Z digest=sha256:7ecf72a8a24f3eccecdd7232da206d5597a4279917e69e14ae849322ab736ef9

Observation c77f0cb6-c974-420b-b245-66d362744f41 · outbound

This paper cites DeepTest: Automated testing of DNN-driven autonomous cars,.

Visual Grounding in Zero-Shot Vision-Language Control DeepTest: Automated testing of DNN-driven autonomous cars,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.544924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.174166Z digest=sha256:8b1555e047309ac6301a56812ae83c47491b5885e6546f66798118b7be1bd00e

Observation 7e56556c-58b4-4bdd-8297-6dc57e25b612 · outbound

This paper cites DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.289522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.263345Z digest=sha256:a07c9f5904cf6fd835c19d348e08815ba91cd6078dce0217814cd3504cbb1d8f

Observation 3d13e243-58ed-4600-a09d-910ea348d10e · outbound

This paper cites Metamorphic testing for semantic invariance in large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing for semantic invariance in large language models,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.128045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.360725Z digest=sha256:097b5bc1cc0772f92219dbc80f64c6af38cdc8e5149a73b171a82b8ec15834eb

Observation b0ec6d09-d4c1-48cf-8a38-2302de9e0bad · outbound

This paper cites Semantic invariance in agentic AI,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic invariance in agentic AI,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.966883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.489405Z digest=sha256:6852b598cc8dc8cf40a9cfcca3cbb5a68a0dae43a66879d95baa4bb82ec261c0

Observation 30cb1895-c93d-4bb4-a201-3ef5de9543b3 · outbound

This paper cites Energy-aware multilingual evaluation of large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Energy-aware multilingual evaluation of large language models,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.818647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.601401Z digest=sha256:aa4fe63c447da4e9da441d11de99b35bd684fd315b371a7a11e999129d3864a2

Observation d6c3c2ca-c79c-444a-b2a7-7605c6b4ab47 · outbound

This paper cites EdgeShard: Efficient LLM inference via collaborative edge computing,.

Visual Grounding in Zero-Shot Vision-Language Control EdgeShard: Efficient LLM inference via collaborative edge computing,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.611251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.719094Z digest=sha256:c35b9aef4a14fb4b901f05a95e0123e80a71b3fd7aaedcb85428c444fd91712a

Observation 225b71da-b3d6-4df4-9682-930d474951cf · outbound

This paper cites Power hungry processing: Watts driving the cost of AI deployment?.

Visual Grounding in Zero-Shot Vision-Language Control Power hungry processing: Watts driving the cost of AI deployment?

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.454238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.811054Z digest=sha256:1d80f63a6786580af824ab8f8bd446e0bfc63370f87f82fbb70c23015579d7cc

Observation 87e4c2cd-8ee4-431e-a0bb-8a28e650ca2b · outbound

This paper cites An environment for autonomous driving decision- making,.

Visual Grounding in Zero-Shot Vision-Language Control An environment for autonomous driving decision- making,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.313176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.906404Z digest=sha256:e368ad8bf6352d10844f007442a5c23b69cb0ee797bdd0f03a4239cfadec7b5e

Observation ff57a3ee-4a1a-4b0c-a83c-a2dc768af9cd · outbound

This paper cites Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,.

Visual Grounding in Zero-Shot Vision-Language Control Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.146672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.001892Z digest=sha256:9ccf0323ad341d7a3291f7db8d191dd3db208b456c0c908bf349f1596bb8f05f

Observation 74bdaa22-0965-492e-9f57-c6556a94b078 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Visual Grounding in Zero-Shot Vision-Language Control Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.105161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.105161Z digest=sha256:db739cd87ce9bca2536a1ad062b605c9ef64e74b15f0e242758ffe3a78175d5e

Observation 7b990891-95c6-4224-a76f-fa9af2c2d912 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models,.

Visual Grounding in Zero-Shot Vision-Language Control Chain-of-thought prompting elicits reasoning in large lan- guage models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.192482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.192482Z digest=sha256:4b585459ae335efe15ea9b89ae66204a280eba3c22948d01199421172eec5e54

Observation e81779ee-264f-462c-bdf0-47724d25e710 · outbound

This paper cites Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,.

Visual Grounding in Zero-Shot Vision-Language Control Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.001615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.303725Z digest=sha256:148ff7e668f415ca7cb392e43431eb558fb079e22aca55fdd309ac18cccba7a5

Observation 4ec7dd12-c123-451e-b0c2-20d72d6cdf58 · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Visual Grounding in Zero-Shot Vision-Language Control Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.374361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.374361Z digest=sha256:025d655b54038e2efb4dd1ee7a3444214790823a888b10fdba6ec1fd6a3686f7

Observation db19dba4-69c2-4470-8fa9-744f8721c355 · outbound

This paper cites Do Prompt-Based Models Really Understand the Meaning of their Prompts?.

Visual Grounding in Zero-Shot Vision-Language Control Do Prompt-Based Models Really Understand the Meaning of their Prompts?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.478970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.478970Z digest=sha256:4ef37dedce87f5c498f8be960b759408ed8470837a5fe21cc67b624626f4d100

Observation d7353651-c3b0-43ac-b1a7-9db0d51c80dc · outbound

This paper cites PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts.

Visual Grounding in Zero-Shot Vision-Language Control PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.549507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.549507Z digest=sha256:df1851a3615bf82af511b518e2da381bc0e4403c1c809264eb881a7d35627e06

Observation 8ebd7ffe-3c10-49de-9ef6-83dd799032d7 · outbound

This paper cites Measuring and improving consistency in pretrained language models,.

Visual Grounding in Zero-Shot Vision-Language Control Measuring and improving consistency in pretrained language models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.854441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.628585Z digest=sha256:567cf0d699bdad1c52fa2fd7e043c58cfe2c298d0224168f6ee42994e4c76770

Observation 0cfbf370-8663-4870-bbf6-caf7a772d722 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Visual Grounding in Zero-Shot Vision-Language Control Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.698424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.698424Z digest=sha256:9388d0431bbe2d89494e38895ba92406e64cc9dd24225140ac7c7cc94727215b

Observation 57408b9b-37d8-41a9-ba60-92e3229aad1c · outbound

This paper cites Probing classifiers: Promises, shortcomings, and ad- vances,.

Visual Grounding in Zero-Shot Vision-Language Control Probing classifiers: Promises, shortcomings, and ad- vances,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.691304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.782731Z digest=sha256:075ef78b08fe22feec79c80935368e1c0aef8be1e02d4522374f186767f70f4d

Observation 3e1d8560-31a1-4c9a-b891-4f3e30f6441a · outbound

This paper cites Con- strained model predictive control: Stability and optimality,.

Visual Grounding in Zero-Shot Vision-Language Control Con- strained model predictive control: Stability and optimality,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.866001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.866001Z digest=sha256:a6890399126282dcd0dd0575bda71658f1ddb3d922e623415cee2b22fc22cc98

Observation bc9e8cc5-c59c-41d8-881c-0040b5fd8e28 · outbound

This paper cites A simple sequentially rejective multiple test procedure,.

Visual Grounding in Zero-Shot Vision-Language Control A simple sequentially rejective multiple test procedure,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.927436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.927436Z digest=sha256:7633ebb8e08f9e69c4ee836c37cba423cb239f03a80756ee366615866273cee8

Observation 0a71f3cc-ccf6-462e-aee2-b81059cc820c · outbound

This paper cites SciPy 1.0: Fundamental algorithms for scientific computing in Python,.

Visual Grounding in Zero-Shot Vision-Language Control SciPy 1.0: Fundamental algorithms for scientific computing in Python,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.541497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:12.001667Z digest=sha256:d3bfa3150eddb2337bff7a11423f97ba3a05eb018b8e04e3fec2f3c1dac23b13

Observation f6350d25-8935-4296-a0fa-a42e0757c57c · outbound

This paper cites Transformers: State-of-the-art natural language pro- cessing,.

Visual Grounding in Zero-Shot Vision-Language Control Transformers: State-of-the-art natural language pro- cessing,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.424487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:12.082156Z digest=sha256:8acb6b0e0f6ee4c3533e8dd3800b8233fbe306f1acd07ebffdbeb9336224433c

Pith citing papers

No inbound Pith citation observations are available.