Pith. sign in

Paper Citation Record · LEDGER

Visual Grounding in Zero-Shot Vision-Language Control

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2608.06154.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06154 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:12.082156Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a603f8c3-022b-4147-8169-63fe57346953 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Visual Grounding in Zero-Shot Vision-Language Control Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.790543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:08.270889Z digest=sha256:ab2fdc434af4deb4d8f146b15f5eb4ac682a16e116175ca5276e69c55f2ee44d

Observation 0e43c405-639a-4c5b-b059-ff170912dad6 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Visual Grounding in Zero-Shot Vision-Language Control Improved Baselines with Visual Instruction Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.301524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.301524Z digest=sha256:f49323cfc153105629d3f6864f35517a45ef96067f94097ad5cbda443783559b

Observation 4603868a-d4e8-4c73-89f7-1d53b5515177 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.373123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.373123Z digest=sha256:e0a7b1512af378c5f956d836ab03dddb96534b6debd954edd9cdc3b9470356e0

Observation a1a210d9-32bb-4b99-8083-5b762cb6b2ef · outbound

This paper cites Qwen2.5-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.442796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.442796Z digest=sha256:487631adf0c485603185dc066c700d6657b71a8f9eaebf8d00898518866ea5b6

Observation dbb8b77a-3727-4987-a03a-e8a06282aa5f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.553499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.553499Z digest=sha256:4671d96f31cbb4b9f71017c65135871520405efa1adbfb51226d256cd5fcf522

Observation fb1a59d1-ed3d-4bad-9ccb-6af984254bd5 · outbound

This paper cites Qwen3-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.643309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.643309Z digest=sha256:6b0b5ae7e3ba7b09a4317fcba0bef4459ff6596ecc41d478bd4e2c3a38ed1fee

Observation 4d55ec77-a452-4a5a-becb-390200596896 · outbound

This paper cites Gemma 4 Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Gemma 4 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.729258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.729258Z digest=sha256:eb53e920d8a8cf980fa6a5ab6544f7f39c1a060b12aa4c24923e55327dda72d9

Observation ec3159d7-5976-4240-90d2-7389098270cd · outbound

This paper cites Qwen3.5: Towards native multimodal agents,.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3.5: Towards native multimodal agents,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.579700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:08.814585Z digest=sha256:f67ccb4499d769508b2a2c15762c71acb06979cce1a0622e5c4a563255c873dc

Observation f585939b-9dd9-4c7b-88d2-52e812ad10b1 · outbound

This paper cites Introducing Mistral 3,.

Visual Grounding in Zero-Shot Vision-Language Control Introducing Mistral 3,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.388390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:08.880443Z digest=sha256:e726101b450debc9cac269b22dbd706b07b05030d63b475aacab4f848e012ec9

Observation 06420827-4300-4bed-81da-c6f12310a09a · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.961126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.961126Z digest=sha256:bdbc6cff147b79e007876077e11fce7865ad58add077a06fe2548837ddc378b6

Observation 9ffb7a4b-5bc3-433f-b84e-657efa32534c · outbound

This paper cites SmolVLM: Redefining small and efficient multimodal models.

Visual Grounding in Zero-Shot Vision-Language Control SmolVLM: Redefining small and efficient multimodal models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.055707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.055707Z digest=sha256:5776f0aa2a5f800ba49c472015c8b154d60459064eae884d9e27be1f4007fa3d

Observation 2e96abb8-7e75-49cf-bad1-44b987d96f15 · outbound

This paper cites NavGPT: Explicit reasoning in vision- and-language navigation with large language models,.

Visual Grounding in Zero-Shot Vision-Language Control NavGPT: Explicit reasoning in vision- and-language navigation with large language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.185974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.150370Z digest=sha256:6323d56938e61d729568000d9d7f07c114e232b839a87d4cb5ca332215009e7d

Observation 2ac4a8d0-ddbe-4845-83ab-753e13619f30 · outbound

This paper cites VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.988861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.191891Z digest=sha256:a8882c297a809f0357ebf9e70205a0d30f2e2e41dff0948de49cbb10ec689285

Observation 1617f2b0-e8b3-4d04-a32d-24e9620e142a · outbound

This paper cites VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.794638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.271864Z digest=sha256:d9f952fd462f1187b502afe122bc495e18eb3d351f471cd5cb15b8a1fc2efc16

Observation f31e68ee-2539-4c67-bb19-0ef0a3340012 · outbound

This paper cites MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,.

Visual Grounding in Zero-Shot Vision-Language Control MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.624658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.376035Z digest=sha256:70dab57e44bffdbbfe63f4b60d8e22153b22d49605ef9316adab7db07bd820a1

Observation 70fc3451-2ab0-43bf-bfdb-0fc46a2249cf · outbound

This paper cites HazardVLM: A video language model for real-time hazard description in automated driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control HazardVLM: A video language model for real-time hazard description in automated driving systems,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.405609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.469350Z digest=sha256:241ba32cca8018f0f34ecadb4e63daa966c08c8cd97ae3662153a272d4be8f34

Observation 25951574-ea08-42cf-87f1-209bb847b839 · outbound

This paper cites LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,.

Visual Grounding in Zero-Shot Vision-Language Control LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.178944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.543585Z digest=sha256:8be7664b8fdd7d4a2edfea6ca843e23520e4dda81dd8ba76f644bf66a50d19ee

Observation c5278006-8a5e-4c69-99f3-2bcd668bf024 · outbound

This paper cites Semantic scene understand- ing with large language models on unmanned aerial vehicles,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic scene understand- ing with large language models on unmanned aerial vehicles,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.962984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.642706Z digest=sha256:7bb4e32bcc747f57e9633b3b27aa8f95e10de347239b3a00ab2f2e537df3c6f7

Observation 91026761-6ef7-4e28-b7b5-ce3798e118e0 · outbound

This paper cites Shortcut learning in deep neural networks,.

Visual Grounding in Zero-Shot Vision-Language Control Shortcut learning in deep neural networks,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.715223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.715223Z digest=sha256:43b1a6c6257133a09f136ca527cd8fe2c52716e741409ffdd30718587c43be92

Observation f29b433f-6864-482b-8456-418ec2180865 · outbound

This paper cites Beyond accuracy: Behavioral testing of NLP models with CheckList,.

Visual Grounding in Zero-Shot Vision-Language Control Beyond accuracy: Behavioral testing of NLP models with CheckList,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.751966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:09.835083Z digest=sha256:ed4e147f4454ae678318b03191d8f2404ccb30a0fb073c0db24a6f9a78346701

Observation ca7c9e58-04a9-4a16-9f56-56cadde12161 · outbound

This paper cites Holistic Evaluation of Language Models.

Visual Grounding in Zero-Shot Vision-Language Control Holistic Evaluation of Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.972585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.972585Z digest=sha256:2244415d17e2c0891844e2df7060d55c944e211a2a86f2a069f7f197fbc61c89

Observation e00c28b0-ce6a-4c24-a854-1d9bf318dca4 · outbound

This paper cites Metamorphic testing: A review of challenges and opportunities,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing: A review of challenges and opportunities,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:10.087742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:10.087742Z digest=sha256:b78ab5a5c608ca256e05bfb5f05732133b6446d1fc881adce7ce0f477b2fc3b6

Observation c77f0cb6-c974-420b-b245-66d362744f41 · outbound

This paper cites DeepTest: Automated testing of DNN-driven autonomous cars,.

Visual Grounding in Zero-Shot Vision-Language Control DeepTest: Automated testing of DNN-driven autonomous cars,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.544924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.174166Z digest=sha256:cf62df0c13d8355761c83dbe9b99ba052a9d29dcc4e444117726b15781038667

Observation 7e56556c-58b4-4bdd-8297-6dc57e25b612 · outbound

This paper cites DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.289522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.263345Z digest=sha256:516891b0ed8c581253c0fe9b0db799f76c4cf945cecec710b2b0e6c0e4ea2638

Observation 3d13e243-58ed-4600-a09d-910ea348d10e · outbound

This paper cites Metamorphic testing for semantic invariance in large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing for semantic invariance in large language models,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.128045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.360725Z digest=sha256:1d1dd85914fc529369cbccfebd1bfd66f7b8fbfc0ccd8d38706891f23a3e2449

Observation b0ec6d09-d4c1-48cf-8a38-2302de9e0bad · outbound

This paper cites Semantic invariance in agentic AI,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic invariance in agentic AI,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.966883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.489405Z digest=sha256:7982684f51db789f29a7e1c0e3c32eb758de33bcd0010782e05bf45e258ffc9d

Observation 30cb1895-c93d-4bb4-a201-3ef5de9543b3 · outbound

This paper cites Energy-aware multilingual evaluation of large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Energy-aware multilingual evaluation of large language models,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.818647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.601401Z digest=sha256:c984a072624728b4b56ba759491a5e09770f5f2ee0e8c36dcea62d4b7bcf414b

Observation d6c3c2ca-c79c-444a-b2a7-7605c6b4ab47 · outbound

This paper cites EdgeShard: Efficient LLM inference via collaborative edge computing,.

Visual Grounding in Zero-Shot Vision-Language Control EdgeShard: Efficient LLM inference via collaborative edge computing,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.611251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.719094Z digest=sha256:b5ff8e84e74d890d7fae2ac17f6cbde0e7118daab6b0ff486617d5065558a324

Observation 225b71da-b3d6-4df4-9682-930d474951cf · outbound

This paper cites Power hungry processing: Watts driving the cost of AI deployment?.

Visual Grounding in Zero-Shot Vision-Language Control Power hungry processing: Watts driving the cost of AI deployment?

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.454238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.811054Z digest=sha256:991169a10d1011ed0f50a3f70a90fb079be8a9a4dbe220188dc987a3ed9eb8e7

Observation 87e4c2cd-8ee4-431e-a0bb-8a28e650ca2b · outbound

This paper cites An environment for autonomous driving decision- making,.

Visual Grounding in Zero-Shot Vision-Language Control An environment for autonomous driving decision- making,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.313176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:10.906404Z digest=sha256:542959ef4fce3fec268fea2cfc14297dade2103fa4a5584df03449a0ec2502f5

Observation ff57a3ee-4a1a-4b0c-a83c-a2dc768af9cd · outbound

This paper cites Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,.

Visual Grounding in Zero-Shot Vision-Language Control Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.146672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.001892Z digest=sha256:092e4e94e591545971dcccf16781af178d7d28069fdad16df32808baf672be64

Observation 74bdaa22-0965-492e-9f57-c6556a94b078 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Visual Grounding in Zero-Shot Vision-Language Control Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.105161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.105161Z digest=sha256:4cc3f65a0aa4dc91240c8712d4483ab854df3d728ee240588bc24a50ec069892

Observation 7b990891-95c6-4224-a76f-fa9af2c2d912 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models,.

Visual Grounding in Zero-Shot Vision-Language Control Chain-of-thought prompting elicits reasoning in large lan- guage models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.192482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.192482Z digest=sha256:2106e73c214aca3297503f79176dd3d57c466717e5a73af8d65cc90fa143a0a0

Observation e81779ee-264f-462c-bdf0-47724d25e710 · outbound

This paper cites Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,.

Visual Grounding in Zero-Shot Vision-Language Control Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.001615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.303725Z digest=sha256:22ae8da7eaf7faeedeed069edf3ecb05e331d174677be3500616a9e53e762cbb

Observation 4ec7dd12-c123-451e-b0c2-20d72d6cdf58 · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Visual Grounding in Zero-Shot Vision-Language Control Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.374361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.374361Z digest=sha256:28b0bd3229096a58d31e7324de74b5fe4b235a20185a5e5b6eb7b5e1e568ecff

Observation db19dba4-69c2-4470-8fa9-744f8721c355 · outbound

This paper cites Do Prompt-Based Models Really Understand the Meaning of their Prompts?.

Visual Grounding in Zero-Shot Vision-Language Control Do Prompt-Based Models Really Understand the Meaning of their Prompts?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.478970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.478970Z digest=sha256:f65cc6f17790fdb570635d66794840b36eed82c5a96730cdea61b54cb207d294

Observation d7353651-c3b0-43ac-b1a7-9db0d51c80dc · outbound

This paper cites PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts.

Visual Grounding in Zero-Shot Vision-Language Control PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.549507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.549507Z digest=sha256:eb51a6e9cfc89cde797db2bbc88479331d28e239b89f115fbbb1950aacd058c3

Observation 8ebd7ffe-3c10-49de-9ef6-83dd799032d7 · outbound

This paper cites Measuring and improving consistency in pretrained language models,.

Visual Grounding in Zero-Shot Vision-Language Control Measuring and improving consistency in pretrained language models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.854441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.628585Z digest=sha256:02ff121088232df2e2f6e7bfb723563c90ae05c1124b5e3e750f0c2a1b34d0a8

Observation 0cfbf370-8663-4870-bbf6-caf7a772d722 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Visual Grounding in Zero-Shot Vision-Language Control Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.698424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.698424Z digest=sha256:3c214cc89ce947595aea778c5d1245cc06b9feb99823c962efc46bc7269dda56

Observation 57408b9b-37d8-41a9-ba60-92e3229aad1c · outbound

This paper cites Probing classifiers: Promises, shortcomings, and ad- vances,.

Visual Grounding in Zero-Shot Vision-Language Control Probing classifiers: Promises, shortcomings, and ad- vances,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.691304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:11.782731Z digest=sha256:2f4e8936cbf7403fc0536e7c1805db615644304083ba6b6c39f19a59565af7a5

Observation 3e1d8560-31a1-4c9a-b891-4f3e30f6441a · outbound

This paper cites Con- strained model predictive control: Stability and optimality,.

Visual Grounding in Zero-Shot Vision-Language Control Con- strained model predictive control: Stability and optimality,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.866001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.866001Z digest=sha256:3553b633773293f9a576935e49e7cb020e599fb2b3b6976617e78667ebd6e2a8

Observation bc9e8cc5-c59c-41d8-881c-0040b5fd8e28 · outbound

This paper cites A simple sequentially rejective multiple test procedure,.

Visual Grounding in Zero-Shot Vision-Language Control A simple sequentially rejective multiple test procedure,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.927436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.927436Z digest=sha256:c1bf36fb6c4d9d475881752e924406f153b4b7b6a46a392dc720122146690dd3

Observation 0a71f3cc-ccf6-462e-aee2-b81059cc820c · outbound

This paper cites SciPy 1.0: Fundamental algorithms for scientific computing in Python,.

Visual Grounding in Zero-Shot Vision-Language Control SciPy 1.0: Fundamental algorithms for scientific computing in Python,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.541497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:12.001667Z digest=sha256:5a6cdb780bfa3617284472023c89143425a8c75fbbb41fa84957c5a5cb4b8f63

Observation f6350d25-8935-4296-a0fa-a42e0757c57c · outbound

This paper cites Transformers: State-of-the-art natural language pro- cessing,.

Visual Grounding in Zero-Shot Vision-Language Control Transformers: State-of-the-art natural language pro- cessing,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.424487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T13:57:12.082156Z digest=sha256:fe87314c176df17c32239502a6f0a36af3f4f9ff9b4d5dca22b30f4350da0ca5

Pith citing papers

No inbound Pith citation observations are available.