Pith. sign in

Paper Citation Record · LEDGER

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

As of 20 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 4 inbound Pith citation observations for arXiv:2505.14071.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14071 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:31.291568Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:45:42.708353Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T23:31:16.269375Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved7
  • parse uncertain4
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fe765d21-e912-4ff4-8f0f-2ef3bce18cb5 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:33.741602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:29.989830Z digest=sha256:fc5e3718458658fc300d3508cd501cd40061119ef87c84644f5bf6bb0c9b173d

Observation 196a1e9a-71b0-4ac1-b608-7f780741b43c · outbound

This paper cites 2 OLMo 2 Furious.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models 2 OLMo 2 Furious

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.875453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:29.875453Z digest=sha256:96fa8ed554dfc1941a9ff9ff38ae5ca5ad42bb57048ad796a0ce7e452fad8469

Observation 46a726e3-8d5e-4299-8b80-e784596effdd · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 3

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.003136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:30.578966Z digest=sha256:e0427b7f4de824a139e5891ecbf5657763f0673e5d555c67f40eea760fcaf3d8

Observation 59fc1de5-1f45-4072-b6a3-10035225c09d · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:33.544136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:30.148065Z digest=sha256:b6ecb0b39f1bc2a149ebeaa9f33d8c025875bcc8ad695f2a9cee8753c6ff715c

Observation e70cef07-cd93-415f-be82-c2b78c35f96a · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 5

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.352863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:30.236996Z digest=sha256:9926f5dbbbf734600fd59263e82f7ba53537e9b9d15376eedc5fa407841c020f

Observation a8625e75-cbaf-4db0-a92e-f31f3a07b8cf · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 6

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:33.191352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:30.410383Z digest=sha256:2ca494d0daa8bab97523477c59699a87c77938515e8275def084acc90e44e3e0

Observation f0c5ad1c-ce68-4f24-8fc9-2e406c1fe6c2 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 8

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T15:42:32.855076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:30.765852Z digest=sha256:edd6448a2117bfddd098e2ed4ded020d6882930f442f771cba8621d132295371

Observation 57c109e9-1f8c-4f8e-a646-79b430dfadd6 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.704442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:30.993716Z digest=sha256:09c2db2825a61826cb01934105c575cc4cf222c14d340dfb74e54dfb49b48df6

Observation 4f9dcd89-fe51-427e-a717-c1ad89b1420d · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.453474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:31.082517Z digest=sha256:df85b431a0c89c33362b175eb9c27f291ea93ad0e433ebe0bc0fc800d3218c85

Observation 90024b35-b02d-4768-8692-0ea109b64263 · outbound

This paper cites an unresolved cited work.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:42:32.136059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:31.188797Z digest=sha256:57a35c4988cdf90ee3818286a7cd43cf9db777901f77ff663dda91cfb1eb2f80

Observation 8ab9e03c-c6f7-4819-8d0d-b6877b6567b3 · outbound

This paper cites INSTRUCTION:.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models INSTRUCTION:

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:42:31.726510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:42:31.291568Z digest=sha256:2dba38d7efa7113dbfd53bb51dad3e98841533cb0f70fde0c06ca0d62939d10e

Observation 1853cc4c-7d87-4f40-9ec0-ad89cb879819 · outbound

This paper cites Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders.

Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:29.746277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:29.746277Z digest=sha256:e59e8575b2e8f4344bf3078a67c70cd0419c54012d8eceb52195f7a4f44c300f

Pith citing papers

Observation e4e68646-ed60-4f23-a0a8-2c927f4350f9 · inbound

Resa: Transparent Reasoning Models via SAEs cites this paper.

Resa: Transparent Reasoning Models via SAEs Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:42.708353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:42.708353Z digest=sha256:ea763565c8d82fc8ef8ce737171575a96fb3f2784774a7a7f72f335fdb996b5c

Observation 2b2fd7ac-43cd-4335-9b36-25aaee936410 · inbound

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models cites this paper.

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T13:45:26.425121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:45:26.425121Z digest=sha256:e1d8ee274c5a001613640d664707bbfcef9f41ca2ccf2a592d7a4cf8731926ad

Observation 6092d87a-c1bf-4fb3-aa3c-e8ea842bf7e5 · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.272701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:3e120282ab91e611dc7d6d195ca7526634879e59e7b0f496d95fbd6527acda43

Observation 80f05176-4831-4e3c-a3bd-126c5f59d985 · inbound

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering cites this paper.

Do Unified Multimodal Models Think in One Space? A Lens Through Cross-Branch Steering Textual Steering Vectors Can Improve Visual Understanding in Multimodal Large Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T16:36:39.333944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:36:39.333944Z digest=sha256:96c361f376567a3e955c528acd80d77a6ed74de73454f413898427c217780942