Pith. sign in

Paper Citation Record · LEDGER

A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2502.17516.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.17516 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T04:27:07.900572Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9bc13ced-8fea-44d0-b73a-58019c5a9ef7 · inbound

LLM-Powered AI Agent Systems and Their Applications in Industry cites this paper.

LLM-Powered AI Agent Systems and Their Applications in Industry A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T14:06:37.994592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T14:05:54.535411Z digest=sha256:27bc289eda0875bfba03dabc5f89a9fd1e5b509ab33ac4160e07660423c8be3f

Observation c22e5c1e-ea0f-4fa7-b0fb-495a22d61f48 · inbound

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models cites this paper.

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:21:36.538648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T16:21:20.463222Z digest=sha256:8465aa0d46f734278cf5a66a598f6bf01a52df67e5dda314d70134093e68ff07

Observation 8b8fabd1-c5e3-438d-995c-ed58af42c9ea · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 187

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:40:54.826576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:0b15242559e02419e51ff13f1d51a7642057c3479a17db13d3b66575441cb038

Observation c1a94952-6fb7-4e74-a193-88657279fc3f · inbound

Wearable AI in the Era of Large Sensor Models cites this paper.

Wearable AI in the Era of Large Sensor Models A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:30:58.075516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T16:35:36.541995Z digest=sha256:eb418a7ecfecc88295f1d2a4ee63b9e9d4c41b4663cec97d39aed0da8bca3d70

Observation 80f20282-2381-4d37-ac40-15f948f5a11c · inbound

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models cites this paper.

Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Models A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:11:53.748635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T07:02:02.752466Z digest=sha256:5b30435d1b312491daa089fbe2afda73010e72e38a068e9119f09b242fa2db58

Observation 9c3af7c1-2df2-444b-ba12-b6abfec22626 · inbound

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models cites this paper.

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:46:10.141459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T05:44:37.891638Z digest=sha256:ee1f1da6dfaca6ba25b61b639f572f61d8278e1f28721edfed0313dfd9656969

Observation f2855f91-1f3f-48db-b801-98d481d39e39 · inbound

SimReg: Achieving Higher Performance in the Pretraining via Embedding Similarity Regularization cites this paper.

SimReg: Achieving Higher Performance in the Pretraining via Embedding Similarity Regularization A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:16:28.897313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T03:33:45.192139Z digest=sha256:dbdcf04d4b2996d391aa885e38e56a549bde0d89bed617ba42f718180b187eb9

Observation 3e6c05ac-1257-45cf-a960-2d6dd9c03c03 · inbound

Do Factual Recall Mechanisms Carry over from Text to Speech in Multimodal Language Models? cites this paper.

Do Factual Recall Mechanisms Carry over from Text to Speech in Multimodal Language Models? A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T06:31:10.545340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T06:26:19.580671Z digest=sha256:623f0d8969a92aa156845aa4e50d8d3fc185401ecd57e40e4aefc481b6921681

Observation 783f212b-fc43-402d-a655-951a5df70ffb · inbound

The physics of AI weather models cites this paper.

The physics of AI weather models A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:25:14.433506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-25T02:20:39.109304Z digest=sha256:c4402f0253a1b8826f192c1eca2ef27bc90d72e864834b9dbc20ad95f887ef7e

Observation e67e3ea1-366c-4208-a4c4-bd8dba2770c4 · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.458106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:e1877d45091600f4cd01e57348e6ae012600273fcc6d36d94129acc67afce897

Observation eb37bca8-edd0-43c5-9bb4-443e7ca2f79c · inbound

Who Wins the Conflict? Mechanistic Interpretability of Text Bias in Audio LLMs cites this paper.

Who Wins the Conflict? Mechanistic Interpretability of Text Bias in Audio LLMs A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:39:24.638089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T19:36:40.901674Z digest=sha256:b2fdb346de7504262c389f06502260f8ef2c2ba4fbb36c1682b5ca9b566097e3

Observation bdd90d0c-74b8-461e-995e-898a784fe6d3 · inbound

FairFlow: Demystifying and Mitigating Stereotype Bias in Text-to-Image Diffusion Transformers cites this paper.

FairFlow: Demystifying and Mitigating Stereotype Bias in Text-to-Image Diffusion Transformers A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T04:19:11.697328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T04:19:11.697328Z digest=sha256:aaaa49bcb6b32c6bb7d7a53cea995c2dbcbeea08427b24550a33a9168ef98121

Observation c4a06452-706b-4a01-866e-ced56ed66d42 · inbound

Unsupervised Features Mining via Activation Geometry cites this paper.

Unsupervised Features Mining via Activation Geometry A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T20:51:42.052454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:51:42.052454Z digest=sha256:12110c018974555578db2606516d3b3d6038307c57e0f5f9c2410cac2bee1b68

Observation 5f34887d-bfc9-4ea3-84a8-1a5900ce8322 · inbound

How Do VLMs Fail? Vision-Operation Misalignment in Compositional VQA cites this paper.

How Do VLMs Fail? Vision-Operation Misalignment in Compositional VQA A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:11.083317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:26:11.083317Z digest=sha256:85a7640b8ea80a36a9d0935f1db37bf74909942f3dbabfb27886c765b4398795

Observation 3198cb64-e105-47b3-a1b1-53f7b87c65d9 · inbound

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs cites this paper.

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T01:33:51.250429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:33:51.250429Z digest=sha256:b428709595ff2b1d6fcd3e0cde0947c8ae301ff76b80c285d899ec50fb38d86b

Observation 5fabadee-bab1-4deb-b1f7-f996f485bf3d · inbound

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs cites this paper.

Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T04:27:07.900572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T04:27:07.900572Z digest=sha256:76b39db8534c7a159c59e75768c7e7dc8f21b6e6377c1fde69edb15297ed6d44

Observation d6e0eebd-6527-463a-817f-2df09474b89b · inbound

Through the LENS: Local Geometric Decomposition of Vision-Language Model Representations cites this paper.

Through the LENS: Local Geometric Decomposition of Vision-Language Model Representations A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T00:46:24.026248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:46:24.026248Z digest=sha256:433647adde67b126f3cff01a479744ee6563317d5ae8017018ed0d2383cc9963