Pith. sign in

Paper Citation Record · LEDGER

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

As of 20 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2509.25533.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.25533 v2

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T13:45:27.971234Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0288cf69-3b9f-47a9-988f-623adeb6abbc · outbound

This paper cites an unresolved cited work.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.971234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.971234Z digest=sha256:a737b55b19f10c4873f460d6945db66c6cde6dd057577c1e371096119a7a476d

Observation b9798a55-6502-4a83-9d66-1fb0777b9aad · outbound

This paper cites How Robust is Google's Bard to Adversarial Image Attacks?.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models How Robust is Google's Bard to Adversarial Image Attacks?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.355350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.355350Z digest=sha256:8b5b00e05053321d8bf15510a22a1fb18223626b4f164cdfe041d1575376eff5

Observation aa488af9-c9f8-4af5-b05b-c8f323acbc04 · outbound

This paper cites X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.695560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.695560Z digest=sha256:2d6f597ee5317f4eade33e0517168f4bec6501d2e9c5e51f3241bb99a83c1157

Observation 55cac5bf-0221-47e5-8ed7-65393e148cbc · outbound

This paper cites Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.820422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.820422Z digest=sha256:c7899b6d884e552ded7148ab11f2ec2ab0892d98e1c48053ec670e762b1b676e

Observation b06be588-7243-45e2-a00d-d19b24ded7b8 · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Steering Llama 2 via Contrastive Activation Addition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.933234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.933234Z digest=sha256:96979eaba72fac61e3ceec38d806862ace5128a2186dda6fdd2f819a290a61eb

Observation 9a351771-dc91-40a1-969e-0b6c8e39baf7 · outbound

This paper cites Visual Adversarial Examples Jailbreak Aligned Large Language Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Visual Adversarial Examples Jailbreak Aligned Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.025943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.025943Z digest=sha256:8dc3c69c0c53175331c3e663232a561baff43ddf48f9f4da7c3400796689efc7

Observation aef1dfde-7d44-4d8d-a5e6-2aa53e5b74fe · outbound

This paper cites Failures to Find Transferable Image Jailbreaks Between Vision-Language Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Failures to Find Transferable Image Jailbreaks Between Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.114819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.114819Z digest=sha256:2aee0e453132019f8bb214b6a7c6b33e6e4d7b7a0bd726d81e0b2d7f284b692c

Observation 54bd8f24-2f21-4381-870a-fa1e1b82fc8b · outbound

This paper cites Jailbreak in pieces: Compositional Adversarial Attacks on Multi-Modal Language Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Jailbreak in pieces: Compositional Adversarial Attacks on Multi-Modal Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.193722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.193722Z digest=sha256:6971d01ae93b725836d10fa9441b0222e45b57a1a2c56ea9ae79daff74298407

Observation ef1ac527-737b-4822-910a-5247718a6f87 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.329648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.329648Z digest=sha256:f6624c3b9400d86d30c3aaf73055d2cd7749e6197fb9be9c4edf34e286f2bfd0

Observation ad59e0bd-557e-4d32-8ad1-01bf5ef03063 · outbound

This paper cites Steering Language Models With Activation Engineering.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Steering Language Models With Activation Engineering

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.475815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.475815Z digest=sha256:0597f411acbc67ef16f2d3ac5176fd7e4c5b882096cf2d49e381dc8305af2e4b

Observation fef9a84f-9bd2-402d-979e-78f556efb848 · outbound

This paper cites AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models AxBench: Steering LLMs? Even Simple Baselines Outperform Sparse Autoencoders

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.597133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.597133Z digest=sha256:96b4ce70eea645a363b51981a79a519380a7484d3bc787b7299273fb4a3337c1

Observation fc3e027f-41fb-48fd-8285-5f1fd1a19a74 · outbound

This paper cites On Evaluating Adversarial Robustness of Large Vision-Language Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models On Evaluating Adversarial Robustness of Large Vision-Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.722293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.722293Z digest=sha256:e1041acc96691672980480f257803362c031be629b5b21d0f1ba9c7dd106e85c

Observation 698190be-f818-430b-bb02-7e330f05392d · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:27.852811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:27.852811Z digest=sha256:4b638841351837c330ac4025fe662addeaad162fce447665eec0838c0e752cbc

Observation 2b2fd7ac-43cd-4335-9b36-25aaee936410 · outbound

This paper cites Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.425121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.425121Z digest=sha256:e1d8ee274c5a001613640d664707bbfcef9f41ca2ccf2a592d7a4cf8731926ad

Observation ebef0b59-8eb2-44e1-8a69-0e328e139d75 · outbound

This paper cites Controlling Large Language Models Through Concept Activation Vectors.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Controlling Large Language Models Through Concept Activation Vectors

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.189146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.189146Z digest=sha256:d157d430b3a015e3129eb6c93d91ae82282028a33593ffc08ad4cce46c49a4a6

Observation c82158bb-6354-4108-8dac-4d49dca84707 · outbound

This paper cites Image Hijacks: Adversarial Images can Control Generative Models at Runtime.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Image Hijacks: Adversarial Images can Control Generative Models at Runtime

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.101151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.101151Z digest=sha256:84646bcaf05573541040d5a21e9c90d307d5fe7391dc752bbf987e7b277c6fcd

Observation 1f4b8a14-b70f-41ac-8c58-79ba983a53fc · outbound

This paper cites Measuring Massive Multitask Language Understanding.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Measuring Massive Multitask Language Understanding

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.549161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.549161Z digest=sha256:88745e9644ec3744c1339177ac3178334f869d051de9b757be4d7736928d0c0b

Pith citing papers

No inbound Pith citation observations are available.