Pith. sign in

Paper Citation Record · LEDGER

Visual Grounding in Zero-Shot Vision-Language Control

As of 10 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2608.06154.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.06154 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:57:12.082156Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy25
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a603f8c3-022b-4147-8169-63fe57346953 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Visual Grounding in Zero-Shot Vision-Language Control Learning transferable visual models from natural language supervision,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.790543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:08.270889Z digest=sha256:d900bd0dfb8ca4b5e4e6ec334db58fa4ca79a2c5e2aadc85c16c6a757b09012d

Observation 0e43c405-639a-4c5b-b059-ff170912dad6 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Visual Grounding in Zero-Shot Vision-Language Control Improved Baselines with Visual Instruction Tuning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.301524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.301524Z digest=sha256:851493dfe1605ed1602fcac0a109328e457bc79c967d821c789edcdb31134e6c

Observation 4603868a-d4e8-4c73-89f7-1d53b5515177 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.373123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.373123Z digest=sha256:953bb09033c6c22cada421f052f41313fbf8317ab2c2fcf8b1eff481f389d2fb

Observation a1a210d9-32bb-4b99-8083-5b762cb6b2ef · outbound

This paper cites Qwen2.5-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.442796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.442796Z digest=sha256:6828b5f866988d5ac1bbc8fe09c2b9e10ff50e08e3568ded33e10352495faac4

Observation dbb8b77a-3727-4987-a03a-e8a06282aa5f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.553499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.553499Z digest=sha256:b449186459fe0328d9784b83615c59301fbc4fb3942c2bcf1e747d4eb8f1af6f

Observation fb1a59d1-ed3d-4bad-9ccb-6af984254bd5 · outbound

This paper cites Qwen3-VL Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.643309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.643309Z digest=sha256:c2d31dfc570acd9b641e8da8119d771e9b00403ebc36cc9e98e71716894e1d08

Observation 4d55ec77-a452-4a5a-becb-390200596896 · outbound

This paper cites Gemma 4 Technical Report.

Visual Grounding in Zero-Shot Vision-Language Control Gemma 4 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.729258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.729258Z digest=sha256:d9af2fad845bbd553b1448f1dbc384be5a658f30f65a304a9009f3ddd853b4f6

Observation ec3159d7-5976-4240-90d2-7389098270cd · outbound

This paper cites Qwen3.5: Towards native multimodal agents,.

Visual Grounding in Zero-Shot Vision-Language Control Qwen3.5: Towards native multimodal agents,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.579700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:08.814585Z digest=sha256:77a5711f0330fdfb586dd5ecfb003f3ae94631396b4b3352829a78784cf50bda

Observation f585939b-9dd9-4c7b-88d2-52e812ad10b1 · outbound

This paper cites Introducing Mistral 3,.

Visual Grounding in Zero-Shot Vision-Language Control Introducing Mistral 3,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.388390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:08.880443Z digest=sha256:8dbf261209480d282ab180bf3d03e5af8307e40242a75dfb27f232752d3bd55c

Observation 06420827-4300-4bed-81da-c6f12310a09a · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

Visual Grounding in Zero-Shot Vision-Language Control MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:08.961126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:08.961126Z digest=sha256:ca42e53b72b7876bca6b16b50adaafcaf16875ca01b94879db27f5743001c019

Observation 9ffb7a4b-5bc3-433f-b84e-657efa32534c · outbound

This paper cites SmolVLM: Redefining small and efficient multimodal models.

Visual Grounding in Zero-Shot Vision-Language Control SmolVLM: Redefining small and efficient multimodal models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.055707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.055707Z digest=sha256:e3671c42193dfe3004b46ca5244599ee2982752adae51f0d65588d600e70f0f8

Observation 2e96abb8-7e75-49cf-bad1-44b987d96f15 · outbound

This paper cites NavGPT: Explicit reasoning in vision- and-language navigation with large language models,.

Visual Grounding in Zero-Shot Vision-Language Control NavGPT: Explicit reasoning in vision- and-language navigation with large language models,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:16.185974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.150370Z digest=sha256:61944a07babbbbe6b3e365daf70874521521542c4e7f5b6bbd9f2f97b4dc0452

Observation 2ac4a8d0-ddbe-4845-83ab-753e13619f30 · outbound

This paper cites VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-Social-Nav: Socially aware robot navigation through scoring using vision-language models,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.988861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.191891Z digest=sha256:6b3de9f18d9c88e5bb9c169a6d95a9be6c2b7f63b9b18d3b908aa5797397ec53

Observation 1617f2b0-e8b3-4d04-a32d-24e9620e142a · outbound

This paper cites VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,.

Visual Grounding in Zero-Shot Vision-Language Control VLM-GroNav: Robot navigation using physically grounded vision-language models in outdoor environments,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.794638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.271864Z digest=sha256:e1e0f6ef28e6f7d55bd7cc2980b3932b83ae453829caa2c7d1f6fcb069834282

Observation f31e68ee-2539-4c67-bb19-0ef0a3340012 · outbound

This paper cites MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,.

Visual Grounding in Zero-Shot Vision-Language Control MapNav: A novel memory representation via annotated semantic maps for VLM-based vision-and-language navigation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.624658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.376035Z digest=sha256:13929658017c521158955aba7c6030d3623735b89b888f6730e55244dffaccff

Observation 70fc3451-2ab0-43bf-bfdb-0fc46a2249cf · outbound

This paper cites HazardVLM: A video language model for real-time hazard description in automated driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control HazardVLM: A video language model for real-time hazard description in automated driving systems,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.405609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.469350Z digest=sha256:111441ed8b26b9184535cc3ed950f2352d1398d8ab597776168a86b61bde06a2

Observation 25951574-ea08-42cf-87f1-209bb847b839 · outbound

This paper cites LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,.

Visual Grounding in Zero-Shot Vision-Language Control LLM-powered cooperative perception framework for mixed UA V-vehicle platoons,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:15.178944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.543585Z digest=sha256:8be68f72b080a4c6780b61d4239448c782414703acaac16afe9776242805cff3

Observation c5278006-8a5e-4c69-99f3-2bcd668bf024 · outbound

This paper cites Semantic scene understand- ing with large language models on unmanned aerial vehicles,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic scene understand- ing with large language models on unmanned aerial vehicles,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.962984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.642706Z digest=sha256:8c295a69b92a54f8c3fa20148b2e053b0d443ee0a317ec9b4d4c050dd1c3c1f2

Observation 91026761-6ef7-4e28-b7b5-ce3798e118e0 · outbound

This paper cites Shortcut learning in deep neural networks,.

Visual Grounding in Zero-Shot Vision-Language Control Shortcut learning in deep neural networks,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.715223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.715223Z digest=sha256:64be0b46a49eeb255160c41ba6efecb073210ffb70805611dd529da191e1f0cb

Observation f29b433f-6864-482b-8456-418ec2180865 · outbound

This paper cites Beyond accuracy: Behavioral testing of NLP models with CheckList,.

Visual Grounding in Zero-Shot Vision-Language Control Beyond accuracy: Behavioral testing of NLP models with CheckList,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.751966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:09.835083Z digest=sha256:c60d4905c01fe1c7031099b4f0d85c21216936a777d797a7c4322db93ba0b330

Observation ca7c9e58-04a9-4a16-9f56-56cadde12161 · outbound

This paper cites Holistic Evaluation of Language Models.

Visual Grounding in Zero-Shot Vision-Language Control Holistic Evaluation of Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:09.972585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:09.972585Z digest=sha256:20ec750c71a5a8c50d2af948c96383f3e511632ddaf362e68e8e847e577ecd85

Observation e00c28b0-ce6a-4c24-a854-1d9bf318dca4 · outbound

This paper cites Metamorphic testing: A review of challenges and opportunities,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing: A review of challenges and opportunities,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:10.087742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:10.087742Z digest=sha256:25d0b31aabdc68b31212445a7181182e9ab8220d43501de0e64c1d1df60cc9fe

Observation c77f0cb6-c974-420b-b245-66d362744f41 · outbound

This paper cites DeepTest: Automated testing of DNN-driven autonomous cars,.

Visual Grounding in Zero-Shot Vision-Language Control DeepTest: Automated testing of DNN-driven autonomous cars,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.544924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.174166Z digest=sha256:94b9a451005b79e9ef0516fc01d617f256e6771981d14d07c3bcbcffc31f9d17

Observation 7e56556c-58b4-4bdd-8297-6dc57e25b612 · outbound

This paper cites DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,.

Visual Grounding in Zero-Shot Vision-Language Control DeepRoad: GAN-based metamorphic testing and input validation framework for autonomous driving systems,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.289522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.263345Z digest=sha256:3dbb9abc87d5ea386e4c225c5ee47407e406c2e82e39d52b718400c7d6cef104

Observation 3d13e243-58ed-4600-a09d-910ea348d10e · outbound

This paper cites Metamorphic testing for semantic invariance in large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Metamorphic testing for semantic invariance in large language models,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:14.128045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.360725Z digest=sha256:98543bcac5dc32df67f9d91a2cfa576313781059fae113c8f63e2b9e624e1c94

Observation b0ec6d09-d4c1-48cf-8a38-2302de9e0bad · outbound

This paper cites Semantic invariance in agentic AI,.

Visual Grounding in Zero-Shot Vision-Language Control Semantic invariance in agentic AI,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.966883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.489405Z digest=sha256:59a2e1a90e1775d508d58073840c1bc6dbc529c3b43dd6f8363f0f17e748d933

Observation 30cb1895-c93d-4bb4-a201-3ef5de9543b3 · outbound

This paper cites Energy-aware multilingual evaluation of large language models,.

Visual Grounding in Zero-Shot Vision-Language Control Energy-aware multilingual evaluation of large language models,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.818647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.601401Z digest=sha256:5afd5de75942612af407ea505f49b13c763e6c7f3a5d0115778646463afd3cef

Observation d6c3c2ca-c79c-444a-b2a7-7605c6b4ab47 · outbound

This paper cites EdgeShard: Efficient LLM inference via collaborative edge computing,.

Visual Grounding in Zero-Shot Vision-Language Control EdgeShard: Efficient LLM inference via collaborative edge computing,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.611251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.719094Z digest=sha256:588749002b8739cd125ec8a2ece0cc85c1c330aa40148f7bd97f3602e61cdfb0

Observation 225b71da-b3d6-4df4-9682-930d474951cf · outbound

This paper cites Power hungry processing: Watts driving the cost of AI deployment?.

Visual Grounding in Zero-Shot Vision-Language Control Power hungry processing: Watts driving the cost of AI deployment?

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.454238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.811054Z digest=sha256:cdb4759d8544fafca0b0dffa85604a317a9b1bd1c42be590284ad81ca0f3cc6e

Observation 87e4c2cd-8ee4-431e-a0bb-8a28e650ca2b · outbound

This paper cites An environment for autonomous driving decision- making,.

Visual Grounding in Zero-Shot Vision-Language Control An environment for autonomous driving decision- making,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.313176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:10.906404Z digest=sha256:6cf9b2bd06acbde7ae19c9c0cc771bb23ea96a60d8860140bc777f9d5b3b6e52

Observation ff57a3ee-4a1a-4b0c-a83c-a2dc768af9cd · outbound

This paper cites Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,.

Visual Grounding in Zero-Shot Vision-Language Control Learning to fly—a Gym environment with PyBullet physics for reinforcement learning of multi-agent quadcopter control,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.146672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:11.001892Z digest=sha256:33a5a888b6132f1c6e465955ebf99731826f667692fe68b58a25776da595afb4

Observation 74bdaa22-0965-492e-9f57-c6556a94b078 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Visual Grounding in Zero-Shot Vision-Language Control Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.105161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.105161Z digest=sha256:0f870c4627a416cb574b24e984553ee64f637802b6a4120f4db79b0114859311

Observation 7b990891-95c6-4224-a76f-fa9af2c2d912 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models,.

Visual Grounding in Zero-Shot Vision-Language Control Chain-of-thought prompting elicits reasoning in large lan- guage models,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.192482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.192482Z digest=sha256:ad06cfcbf9cef7a9741429596d4b0f2698115039152e417f696bb3af6e39ccb8

Observation e81779ee-264f-462c-bdf0-47724d25e710 · outbound

This paper cites Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,.

Visual Grounding in Zero-Shot Vision-Language Control Language models don’t always say what they think: Unfaithful explanations in chain- of-thought prompting,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:13.001615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:11.303725Z digest=sha256:7579079797999059dd2927996d03374ec2473393a6e772e1f76edf9d23b102f6

Observation 4ec7dd12-c123-451e-b0c2-20d72d6cdf58 · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Visual Grounding in Zero-Shot Vision-Language Control Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.374361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.374361Z digest=sha256:3ef41c8aaa685235539d967c809ac88bdea42aa0750684a71aa92b9805d94662

Observation db19dba4-69c2-4470-8fa9-744f8721c355 · outbound

This paper cites Do Prompt-Based Models Really Understand the Meaning of their Prompts?.

Visual Grounding in Zero-Shot Vision-Language Control Do Prompt-Based Models Really Understand the Meaning of their Prompts?

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.478970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.478970Z digest=sha256:47d2d1c7381decdebd323707a55067a201d7dd87c7dee83528cd75c084509e21

Observation d7353651-c3b0-43ac-b1a7-9db0d51c80dc · outbound

This paper cites PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts.

Visual Grounding in Zero-Shot Vision-Language Control PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.549507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.549507Z digest=sha256:e71f8728f03e8abcb3880464dd1a8d8a27b7cc24f3f0d867e70394cc2ea1eee9

Observation 8ebd7ffe-3c10-49de-9ef6-83dd799032d7 · outbound

This paper cites Measuring and improving consistency in pretrained language models,.

Visual Grounding in Zero-Shot Vision-Language Control Measuring and improving consistency in pretrained language models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.854441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:11.628585Z digest=sha256:b6256fe9ecdf97adca698df9c7ebcee5e856bb3534fbe362a10ff1b34daacfe3

Observation 0cfbf370-8663-4870-bbf6-caf7a772d722 · outbound

This paper cites Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity.

Visual Grounding in Zero-Shot Vision-Language Control Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order Sensitivity

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.698424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.698424Z digest=sha256:12d08a5089b997f9a3f39446bbf932f96ccd76e99b2bc99841c8260effd58331

Observation 57408b9b-37d8-41a9-ba60-92e3229aad1c · outbound

This paper cites Probing classifiers: Promises, shortcomings, and ad- vances,.

Visual Grounding in Zero-Shot Vision-Language Control Probing classifiers: Promises, shortcomings, and ad- vances,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.691304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:11.782731Z digest=sha256:5ad1e9ce2ac028b9dd76b63b7b8a9dbd40e1d7eb28431859f326e798e9426079

Observation 3e1d8560-31a1-4c9a-b891-4f3e30f6441a · outbound

This paper cites Con- strained model predictive control: Stability and optimality,.

Visual Grounding in Zero-Shot Vision-Language Control Con- strained model predictive control: Stability and optimality,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.866001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.866001Z digest=sha256:b45fcfa551d3c742a19933f1776f72b23cc7d16c2bbe9bc3349f93447e58dce8

Observation bc9e8cc5-c59c-41d8-881c-0040b5fd8e28 · outbound

This paper cites A simple sequentially rejective multiple test procedure,.

Visual Grounding in Zero-Shot Vision-Language Control A simple sequentially rejective multiple test procedure,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:11.927436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:11.927436Z digest=sha256:97d7425dbcd5fd2d82d1b13a65cb58ca7813210d24e77b8aa7c734f734858682

Observation 0a71f3cc-ccf6-462e-aee2-b81059cc820c · outbound

This paper cites SciPy 1.0: Fundamental algorithms for scientific computing in Python,.

Visual Grounding in Zero-Shot Vision-Language Control SciPy 1.0: Fundamental algorithms for scientific computing in Python,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.541497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:12.001667Z digest=sha256:50b89a93bc2d0ec01b777251b0640bee3f02d4ae8dc5e65965607f406f67080f

Observation f6350d25-8935-4296-a0fa-a42e0757c57c · outbound

This paper cites Transformers: State-of-the-art natural language pro- cessing,.

Visual Grounding in Zero-Shot Vision-Language Control Transformers: State-of-the-art natural language pro- cessing,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:57:12.424487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T13:57:12.082156Z digest=sha256:c3e30a0c79014f51ce847859dd2b1c5d94d8c297a2affef52576069e2a92580d

Pith citing papers

No inbound Pith citation observations are available.