Pith. sign in

Paper Citation Record · LEDGER

Multimodal Model Diffing for Feature Discovery and Control

As of 11 August 2026, this Paper Citation Record lists 99 of 99 outbound references and 0 inbound Pith citation observations for arXiv:2608.09928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09928 v1

Coverage vector

measured 99 of 99 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:17:56.224644Z

measured 99 of 99 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

99 of 99 outbound references displayed

  • verified exact2
  • verified fuzzy22
  • unresolved74
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 63a68806-aeda-41aa-8974-f329e2ebba67 · outbound

This paper cites Pixtral 12b: A new frontier in image and text understanding.

Multimodal Model Diffing for Feature Discovery and Control Pixtral 12b: A new frontier in image and text understanding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.734741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.734741Z digest=sha256:91e7ce18739c75785eaf2b864ce1040e6458a1fa66eab2fc3f6044e1f27ba11d

Observation e980897a-63ad-4ab4-9cae-ab1bace9e2e9 · outbound

This paper cites Golden gate Claude.

Multimodal Model Diffing for Feature Discovery and Control Golden gate Claude

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.740439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.740439Z digest=sha256:3316eb211df822e23cb720505d1b87fc5a5cf1007ef1e1144a70a7d795d5476d

Observation b478f42b-d307-42c5-b16b-5954735aa5a1 · outbound

This paper cites SAE on activation differences.

Multimodal Model Diffing for Feature Discovery and Control SAE on activation differences

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.744964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.744964Z digest=sha256:7ea0b4cc79ac2839eae8b2970831b6afb2bc9becf77f1e6ff708d9e8b2989fa8

Observation eed6e3f5-9874-4c87-b480-11f0d9fdb4ad · outbound

This paper cites Refusal in Language Models Is Mediated by a Single Direction.

Multimodal Model Diffing for Feature Discovery and Control Refusal in Language Models Is Mediated by a Single Direction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.749442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.749442Z digest=sha256:5a801478350eeccd8686749c36806a095511cf9c54d82925afe8c347a467f57c

Observation 15f9a23f-fb05-4140-b13b-89fcd2bd4712 · outbound

This paper cites Revisiting model stitching to compare neural representations.Advances in neural information processing systems, 34:225–236, 2021.

Multimodal Model Diffing for Feature Discovery and Control Revisiting model stitching to compare neural representations.Advances in neural information processing systems, 34:225–236, 2021

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.754302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.754302Z digest=sha256:cb0b39637989d7b708009021248a77c039d1a5544baf4a345e768284ee23586e

Observation 99bad67a-8f11-4798-a8d3-1e4ee9868bcc · outbound

This paper cites Representation Topology Divergence: A Method for Comparing Neural Network Representations.

Multimodal Model Diffing for Feature Discovery and Control Representation Topology Divergence: A Method for Comparing Neural Network Representations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.758852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.758852Z digest=sha256:ea83967e56da72b52ef8b4d29f2f2f5d02db072ac45a4e04eb0b8d6f25f239ca

Observation fb3300c3-6ac8-4870-bef9-0d7a1e819f96 · outbound

This paper cites Understanding information storage and transfer in multi-modal large language models.Advances in Neural Information Processing Systems, 37:7400–7426, 2024.

Multimodal Model Diffing for Feature Discovery and Control Understanding information storage and transfer in multi-modal large language models.Advances in Neural Information Processing Systems, 37:7400–7426, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.764254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.764254Z digest=sha256:262386cd187790b4d41689ad7c14bd8b2b442d8d99cd228652ae83a88fcb4fdd

Observation 775b35d7-5787-42c1-b28c-727175dec30e · outbound

This paper cites Towards monosemanticity: Decomposing language models with dictionary learning.Transformer Circuits Thread, 2023.

Multimodal Model Diffing for Feature Discovery and Control Towards monosemanticity: Decomposing language models with dictionary learning.Transformer Circuits Thread, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.768697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.768697Z digest=sha256:db8a4e38103efe12645bc06f2e32ce5a6b8f192720ccbe6093bdc08b4606b7d5

Observation bfb2af8f-ac38-41fa-a848-228a761cca3a · outbound

This paper cites Stage-wise model diffing.

Multimodal Model Diffing for Feature Discovery and Control Stage-wise model diffing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.773486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.773486Z digest=sha256:4d894270c0c75d8c291ca245693738e824b82dcc7af7ff5eb00b3272660f95d8

Observation 96afae2d-fab5-435d-bf48-229175c2211e · outbound

This paper cites Observing and controlling features in vision-language-action models.arXiv preprint arXiv:2603.05487, 2026.

Multimodal Model Diffing for Feature Discovery and Control Observing and controlling features in vision-language-action models.arXiv preprint arXiv:2603.05487, 2026

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.778192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.778192Z digest=sha256:40667ece1c409cb09a25071aabe1840f2d462b2a16b16c143bf3efb33d9d73f0

Observation bae5706b-c4c3-4713-829f-3cfd8479df19 · outbound

This paper cites Improving Steering Vectors by Targeting Sparse Autoencoder Features.

Multimodal Model Diffing for Feature Discovery and Control Improving Steering Vectors by Targeting Sparse Autoencoder Features

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.782849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.782849Z digest=sha256:8a37bb063a6dfda38fdb341de25f1d58744db71c91ec58696459efa65448a1b3

Observation 21aef9b8-39d4-4a4a-bb5e-514a373a0a0d · outbound

This paper cites Pappas, Florian Tramer, Hamed Hassani, and Eric Wong.

Multimodal Model Diffing for Feature Discovery and Control Pappas, Florian Tramer, Hamed Hassani, and Eric Wong

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.788114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.788114Z digest=sha256:3f47b6880198a97296f7daf71b80f0cb9f7ead5296175ecfa005e16f2303014c

Observation 943068ae-9db0-4d97-a492-eea47ed36f53 · outbound

This paper cites Interpreting and Controlling Vision Foundation Models via Text Explanations.

Multimodal Model Diffing for Feature Discovery and Control Interpreting and Controlling Vision Foundation Models via Text Explanations

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.793237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.793237Z digest=sha256:8f1f20878df44a569875386ddaba743ca634c46c3e21f3b23ca6d851ec1d7a4b

Observation 729e283c-fe58-406d-a371-03d8e94288e4 · outbound

This paper cites LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning.

Multimodal Model Diffing for Feature Discovery and Control LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.798668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.798668Z digest=sha256:0b150850ed780667f5543932b4a5510c05f082cb910212ad239486658c290253

Observation 47c9ed4e-1705-4432-9bf6-b809e4061e02 · outbound

This paper cites Explaining How Visual, Textual and Multimodal Encoders Share Concepts.

Multimodal Model Diffing for Feature Discovery and Control Explaining How Visual, Textual and Multimodal Encoders Share Concepts

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T04:17:57.283192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:55.803988Z digest=sha256:13cef1d8574598c974b0b99e7dc4303a289f763d53a56c7233edaed801c9df9f

Observation 9ea49434-9b1f-4ea9-b45e-0b22dc3bbf00 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

Multimodal Model Diffing for Feature Discovery and Control Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.809262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.809262Z digest=sha256:b48a8c43a50632c1c740b0ad996bf712b074cc7742b3b6a3009fcbcd972f6588

Observation ef347515-df97-4217-990c-bac9c5c5f962 · outbound

This paper cites Case study: Interpreting, manipulating, and controlling CLIP with sparse autoencoders.

Multimodal Model Diffing for Feature Discovery and Control Case study: Interpreting, manipulating, and controlling CLIP with sparse autoencoders

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.814349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.814349Z digest=sha256:58b355e14268dbbeeedecd0e6279cb07d1b833ebbb194253b88d0826f96d35ab

Observation c03b7f1f-509a-432a-bf1a-565eab037242 · outbound

This paper cites Toy Models of Superposition.

Multimodal Model Diffing for Feature Discovery and Control Toy Models of Superposition

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.819230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.819230Z digest=sha256:deed62bc3aa8b08d54b032c9b2917fe060a8bd2b9fe44f33bc0b239f82a6ab22

Observation 907a2509-8501-4d4e-b1ce-4313ef05a681 · outbound

This paper cites Why does unsupervised pre-training help deep learning? 11:625–660, March.

Multimodal Model Diffing for Feature Discovery and Control Why does unsupervised pre-training help deep learning? 11:625–660, March

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.824530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.824530Z digest=sha256:05466af74c67400479f2b0a483d83ba34e4c279bbbe4e5012d2806f5ed18e899

Observation 5dc90ba2-6e4d-4bbf-98b4-ace19979d697 · outbound

This paper cites Interpreting CLIP's Image Representation via Text-Based Decomposition.

Multimodal Model Diffing for Feature Discovery and Control Interpreting CLIP's Image Representation via Text-Based Decomposition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.829626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.829626Z digest=sha256:cfb04a88468092e8108f72482efd35258ab50286bc855a1501b2dfcc2e0355fe

Observation 4227ee35-1ef9-4683-9e9f-b03e81994704 · outbound

This paper cites Scaling and evaluating sparse autoencoders.

Multimodal Model Diffing for Feature Discovery and Control Scaling and evaluating sparse autoencoders

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.834643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.834643Z digest=sha256:448f40dbc2d63cfa25db371defbe0a6fef1ad2da2642bcee13f19d073059f9d4

Observation a67d0a03-cb2c-48bc-99b6-5077d2ca94a8 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Multimodal Model Diffing for Feature Discovery and Control Gemma 2: Improving Open Language Models at a Practical Size

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.839825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.839825Z digest=sha256:94bbcf217d4512401897b7bc504ea17d0f5b4ad5d9ec1edb92a7600525a17313

Observation 2891bc5a-3300-4926-b5a2-2e673fc4e1d2 · outbound

This paper cites FigStep: Jailbreaking large vision-language models via typographic visual prompts.

Multimodal Model Diffing for Feature Discovery and Control FigStep: Jailbreaking large vision-language models via typographic visual prompts

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.845475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.845475Z digest=sha256:c0d43b05d4aac72d0351104c9eb05fe5f5ced79e38ec732716d3f4f66613109b

Observation eb7628dc-8244-471c-97d8-776e74cf4700 · outbound

This paper cites Making the v in vqa matter: Elevating the role of image understanding in visual question answering.

Multimodal Model Diffing for Feature Discovery and Control Making the v in vqa matter: Elevating the role of image understanding in visual question answering

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.850637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.850637Z digest=sha256:cb9c61486b8b028ec4fb629937bfb62bfc87dea275b93c236732044fd07d3674

Observation 383f8bd4-0f3d-4f1e-8663-f5c8552e8985 · outbound

This paper cites Not all features are created equal: A mechanistic study of vision-language-action models.

Multimodal Model Diffing for Feature Discovery and Control Not all features are created equal: A mechanistic study of vision-language-action models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.856252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.856252Z digest=sha256:e4c41c501293de40ec68ce6e105009d254f9ee6d3cbcf01e4c0fc4aa87b6653e

Observation 594550ea-48d5-4b18-a375-73064b1a6c2e · outbound

This paper cites The Llama 3 Herd of Models.

Multimodal Model Diffing for Feature Discovery and Control The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.861229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.861229Z digest=sha256:0190d55945ec8a43718e199be1b7737d69175697e9fd68d773db8bfa5fbb5fc6

Observation 51b802d7-af4b-4cde-979e-97fc6d7c7e49 · outbound

This paper cites Mechanistic interpretability for steering vision-language-action models.

Multimodal Model Diffing for Feature Discovery and Control Mechanistic interpretability for steering vision-language-action models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.865869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.865869Z digest=sha256:0bb7c655290a7d74be15ac7554fbcf3042c57822996348f9b7d6789e7d78ee51

Observation 3b640386-0590-4017-8b10-8dcab1ed9d2b · outbound

This paper cites Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders.

Multimodal Model Diffing for Feature Discovery and Control Llama Scope: Extracting Millions of Features from Llama-3.1-8B with Sparse Autoencoders

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.870806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.870806Z digest=sha256:d1b157438a00f7fe122d1bc19eb5b7c135174de23c63e46a3a8a9db9846742d8

Observation e94666bd-7d15-4980-b0a4-2c789fe46bb3 · outbound

This paper cites In-context learning creates task vectors.

Multimodal Model Diffing for Feature Discovery and Control In-context learning creates task vectors

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.875585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.875585Z digest=sha256:5d0c975ae3d4cc1de6a5bbae6a6ccb441ed24c68436851c15678329fda25f9d4

Observation ed622c66-b2d4-4498-8016-67dc19488d7e · outbound

This paper cites VLSBench: Unveiling Visual Leakage in Multimodal Safety.

Multimodal Model Diffing for Feature Discovery and Control VLSBench: Unveiling Visual Leakage in Multimodal Safety

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.880064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.880064Z digest=sha256:0c78f863df0b7cbe29cd6467c565ba96e66e6c4f5c8440cd842147104b200ee4

Observation 92e3a5d3-924f-4d22-961c-de23b2368db0 · outbound

This paper cites Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training.

Multimodal Model Diffing for Feature Discovery and Control Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.885168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.885168Z digest=sha256:8036587d2b183146d43060be6b4b3a7698cfa34bc75cba7105a3fe478147de42

Observation 60bfe481-7159-4238-b912-6accdf089da1 · outbound

This paper cites Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations.

Multimodal Model Diffing for Feature Discovery and Control Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.889674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.889674Z digest=sha256:ae9068907294249e9b7de55bfd061ea2d7581686cdef0e607c5adef821012f94

Observation dd219166-8d27-475a-a1c4-06ce11c66684 · outbound

This paper cites A “diff” tool for AI: Finding behavioral differences in new models.

Multimodal Model Diffing for Feature Discovery and Control A “diff” tool for AI: Finding behavioral differences in new models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.894475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.894475Z digest=sha256:bf393f465782a1089c320c5ed2f406f3cd82ab89201caf1221ce65209cff6425

Observation 7bdee4c9-2165-49f5-a2c8-8a114e2a8644 · outbound

This paper cites Bridging the VLM and mech interp communities for multimodal interpretability.

Multimodal Model Diffing for Feature Discovery and Control Bridging the VLM and mech interp communities for multimodal interpretability

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.900408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.900408Z digest=sha256:544e30cf78a0468823ee3213dc54a4b6c0522a33c6a0da5a3f452375cf981120

Observation 0a1b3935-db61-4b7d-b663-d9f6cdf18b9d · outbound

This paper cites Steering CLIP's vision transformer with sparse autoencoders.

Multimodal Model Diffing for Feature Discovery and Control Steering CLIP's vision transformer with sparse autoencoders

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.905911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.905911Z digest=sha256:bd42806ca9c0ca522ceb80362646f77020859adbab9cfaf311a363aaa414e3e7

Observation 02d7cbe7-c662-4f20-a2f1-8f252dca6f28 · outbound

This paper cites Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video.

Multimodal Model Diffing for Feature Discovery and Control Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.911157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.911157Z digest=sha256:ba53edd738786dc3e44a2a06cc4678c2025a9bc301618c5679de920fb7f409d8

Observation 4e56b9bf-2caf-4a64-8b63-4fffadd4fa48 · outbound

This paper cites Analyzing Finetuning Representation Shift for Multimodal LLMs Steering.

Multimodal Model Diffing for Feature Discovery and Control Analyzing Finetuning Representation Shift for Multimodal LLMs Steering

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.916191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.916191Z digest=sha256:3421c37076586f98a002eb8a05093b7acf35144bb264da6e715112fe687bf864

Observation bb729be5-e31f-4c6f-95d7-d9726ae0138e · outbound

This paper cites Saes (usually) transfer between base and chat models.

Multimodal Model Diffing for Feature Discovery and Control Saes (usually) transfer between base and chat models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.921070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.921070Z digest=sha256:8e8d3c2cf6e15ddebb143ffa10375879b324c19822979e0776cc9c16998df9ad

Observation 16f919f1-5867-4ed1-b1a8-49dea9800dbf · outbound

This paper cites Similarity of neural network representations revisited.

Multimodal Model Diffing for Feature Discovery and Control Similarity of neural network representations revisited

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.926358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.926358Z digest=sha256:afee48eebe78d0f64ba2d6ab790397cf077a73cbacde27af094b1f9005194938

Observation 07a5bec2-b757-4592-8d0b-4eccf85b6d89 · outbound

This paper cites Sakla, and Kowshik Thopalli.

Multimodal Model Diffing for Feature Discovery and Control Sakla, and Kowshik Thopalli

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.931559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.931559Z digest=sha256:0ae0baf07cefc2c26713714cb1ee15920fac14dc423234ae1dfa734a03bafdd4

Observation 62157067-b288-4ad7-97cb-c5b4e1633c98 · outbound

This paper cites Understanding image representations by measuring their equivariance and equivalence.

Multimodal Model Diffing for Feature Discovery and Control Understanding image representations by measuring their equivariance and equivalence

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.936739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.936739Z digest=sha256:bbebc96fb3c2c156feb5c5d8a45ba62bb0b24e671d1bd3401a0a252973f0d3cc

Observation ef4f6fd3-9c05-4ffc-8be7-8a071882afe6 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

Multimodal Model Diffing for Feature Discovery and Control LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.941614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.941614Z digest=sha256:4f1de215b0a6dd08b902a8356eaf9d0087c92a7c6c460acda3f4c50f1e1ea5fa

Observation ae34dab8-af93-4449-8e3a-061503ffd727 · outbound

This paper cites Inference- time intervention: Eliciting truthful answers from a language model.

Multimodal Model Diffing for Feature Discovery and Control Inference- time intervention: Eliciting truthful answers from a language model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.946783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.946783Z digest=sha256:cff600f321db5f7c581d5593615305d19dd9d7253ee400fe9279477cae73eee2

Observation f66078d7-2322-48be-8ac4-5c1c808c4191 · outbound

This paper cites Images are Achilles’ heel of alignment: Exploiting visual vulnerabilities for jailbreaking multimodal large language models.

Multimodal Model Diffing for Feature Discovery and Control Images are Achilles’ heel of alignment: Exploiting visual vulnerabilities for jailbreaking multimodal large language models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.969271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:55.951510Z digest=sha256:a45b7b1ab683f41ff41fcd20cadd72e0fc94e3665be7326567a9544ad7e236c1

Observation 7c5cf49e-69d0-46fc-b00c-d062e287ae2c · outbound

This paper cites Convergent Learning: Do different neural networks learn the same representations?.

Multimodal Model Diffing for Feature Discovery and Control Convergent Learning: Do different neural networks learn the same representations?

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.956223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.956223Z digest=sha256:79db7004b75da247995310381d1794c898fdcb9992d9fc237e81d53bdb590561

Observation b362ad44-4a2b-4a9d-9394-b8618f4cc42d · outbound

This paper cites Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2.

Multimodal Model Diffing for Feature Discovery and Control Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.961255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.961255Z digest=sha256:2e00ac1533ad64c9aa59715ebe31303e0bf07c9582d20e88baf9110611d21eb7

Observation f5eddaa7-8d8b-40f1-9231-425afcca5f9e · outbound

This paper cites Sparse autoencoders reveal selective remapping of visual concepts during adaptation.

Multimodal Model Diffing for Feature Discovery and Control Sparse autoencoders reveal selective remapping of visual concepts during adaptation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.967019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.967019Z digest=sha256:0fd1e41a2352c26a5596e713635276f237152c771510096ce2ad411ef43f2451

Observation 74cb08c5-3e38-4f66-b5a3-cf3e6d9061f1 · outbound

This paper cites A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models.

Multimodal Model Diffing for Feature Discovery and Control A Survey on Mechanistic Interpretability for Multi-Modal Foundation Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.972218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.972218Z digest=sha256:6a926819c6f7929f392e6370550d605ed0d962f0d8edb997b9e439821eeb2c20

Observation 50c39459-8359-4d8a-8f3b-06a5764fccdb · outbound

This paper cites Sparse crosscoders for cross-layer features and model diffing, October 25 2024.

Multimodal Model Diffing for Feature Discovery and Control Sparse crosscoders for cross-layer features and model diffing, October 25 2024

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.952495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:55.977277Z digest=sha256:618ed08ad089fa96cb20e907203d318d6baa4c0d162a81097dc5ba22941178ae

Observation e242e417-d6d1-466c-8c67-75d87eb17f48 · outbound

This paper cites Visual spatial reasoning.Transactions of the Association for Computational Linguistics, 2023.

Multimodal Model Diffing for Feature Discovery and Control Visual spatial reasoning.Transactions of the Association for Computational Linguistics, 2023

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.934987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:55.981926Z digest=sha256:cfddb0f06499248ceab7eca4941d4bc890fc64e337bc93df2a0d5dfcbdd58408

Observation 428d81df-84f7-4297-95c6-84ba78dfca31 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

Multimodal Model Diffing for Feature Discovery and Control Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.986160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.986160Z digest=sha256:62a9b6c02f561418a5ba53426432caed6586be90685bf4541e7dfcfd3959599e

Observation 159ce224-eb30-4005-bd03-3e7034b33e31 · outbound

This paper cites Improved baselines with visual instruction tuning.

Multimodal Model Diffing for Feature Discovery and Control Improved baselines with visual instruction tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.990284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.990284Z digest=sha256:e82fa937323af48122079822fc3403b104bee073f0312371d3dd483d1b506de6

Observation a8f7a635-04af-41c1-9058-446d263e58fb · outbound

This paper cites MM-SafetyBench: A benchmark for safety evaluation of multimodal large language models.

Multimodal Model Diffing for Feature Discovery and Control MM-SafetyBench: A benchmark for safety evaluation of multimodal large language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.897139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:55.994689Z digest=sha256:3f934698c23888c8402baffe48458bd8345113ba2077c116e98116081e6262f0

Observation 4cc32df0-f6f3-4a06-9080-8ad98d0fa6c4 · outbound

This paper cites OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models.

Multimodal Model Diffing for Feature Discovery and Control OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.999133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.999133Z digest=sha256:37181ae21cfa2ea588a5703bd1b4a383cb61642c15a30a9e7cbd4fcc12faa17f

Observation 2db28554-8e14-41be-8bc8-fc9bd7333573 · outbound

This paper cites Michaud, Yonatan Belinkov, David Bau, and Aaron Mueller.

Multimodal Model Diffing for Feature Discovery and Control Michaud, Yonatan Belinkov, David Bau, and Aaron Mueller

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.880397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.003768Z digest=sha256:bdc9e280147796d2570213ce4c3a2e51b1a0347397a2a0e3629dfbebf61603d3

Observation 1012ade3-0d1d-4240-a17c-094795998f39 · outbound

This paper cites Locating and editing factual associations in GPT.

Multimodal Model Diffing for Feature Discovery and Control Locating and editing factual associations in GPT

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.008128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.008128Z digest=sha256:934e55330e05e0a7b084d52ab98489a30f11cafca4b853dbaa670b93ecd143ca

Observation 464d7b55-d456-457f-86c9-7b103017259e · outbound

This paper cites Robustly identifying concepts introduced during chat fine-tuning using crosscoders.arXiv preprint arXiv:2504.02922, 2025.

Multimodal Model Diffing for Feature Discovery and Control Robustly identifying concepts introduced during chat fine-tuning using crosscoders.arXiv preprint arXiv:2504.02922, 2025

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.012960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.012960Z digest=sha256:4e35f408c6e4d6be1709aff4f8406dc3026dc868f1599d392e84fc382038ca92

Observation cc5bc455-3eba-42e8-901f-b572d332ca4a · outbound

This paper cites What we learned trying to diff base and chat models (and why it matters).LessWrong, 2025.

Multimodal Model Diffing for Feature Discovery and Control What we learned trying to diff base and chat models (and why it matters).LessWrong, 2025

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.853986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.017719Z digest=sha256:1b46e9da01dc716de91dd3017ad94e0bbd6be1e1d14b92762f21c3389305ab7e

Observation 8451658b-066d-4558-848c-483cfff8b398 · outbound

This paper cites Insights on crosscoder model diffing.

Multimodal Model Diffing for Feature Discovery and Control Insights on crosscoder model diffing

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.837250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.022501Z digest=sha256:c4c5c6f28597b06d45c25e4a281ec5634ed2759fc336318b198f40b308c6e6bb

Observation 5db2cfa0-3178-48d0-9c2f-245259e280e8 · outbound

This paper cites Attribution patching: Activation patching at industrial scale.

Multimodal Model Diffing for Feature Discovery and Control Attribution patching: Activation patching at industrial scale

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.818307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.027425Z digest=sha256:25080404855f7c314911c6105c9f0cd07ab6e63b78f3c8d3bf695385e0079dd2

Observation cdd46c79-02f1-4c4e-86bb-b856ba020424 · outbound

This paper cites Towards Interpreting Visual Information Processing in Vision-Language Models.

Multimodal Model Diffing for Feature Discovery and Control Towards Interpreting Visual Information Processing in Vision-Language Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.032176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.032176Z digest=sha256:2b072db814c7749be34f586c1635574f7b4f5c5367b35685cd0a955f3082955f

Observation 623a2370-84b8-4f4a-bf0b-0b8c4a7bd95c · outbound

This paper cites Steering Language Model Refusal with Sparse Autoencoders.

Multimodal Model Diffing for Feature Discovery and Control Steering Language Model Refusal with Sparse Autoencoders

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.038373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.038373Z digest=sha256:80325ed3c9f059ec0bdeda53a0bd35c064608970c43b1c0b878b22beecd6cbfc

Observation bbc722ed-96ce-4a69-89d5-624cf509d697 · outbound

This paper cites Zoom in: An introduction to circuits.Distill, 2020.

Multimodal Model Diffing for Feature Discovery and Control Zoom in: An introduction to circuits.Distill, 2020

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.044310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.044310Z digest=sha256:71a1a9a28464e849ecbe9e4b94cd9b79dcea59025bb4a796b9adfc19acbbc32d

Observation 058d2038-410f-42cd-9c2b-798bf255d421 · outbound

This paper cites Visualizing representations: Deep learning and human beings.

Multimodal Model Diffing for Feature Discovery and Control Visualizing representations: Deep learning and human beings

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.798542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.049202Z digest=sha256:a7539c4cff2b25c1d32b4ced6fd3d2bba3dd0c221eee05c4249ef4413ee3c731

Observation 60c5603e-ea1d-40e2-b130-362652456e2b · outbound

This paper cites Probing the representational power of sparse autoencoders in vision models.

Multimodal Model Diffing for Feature Discovery and Control Probing the representational power of sparse autoencoders in vision models

Reference 65

Resolution
verified exact
raw_fallback, observed 2026-08-11T04:17:56.747243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.054083Z digest=sha256:8e8eb23c0b7de5b0952ccb654d7297663e06951a69d339d62c4e512392be39cc

Observation 5b9793e9-4c4a-4e7c-862e-920e5d638bbe · outbound

This paper cites Gpt-4o-mini: Advancing cost-efficient intelligence.

Multimodal Model Diffing for Feature Discovery and Control Gpt-4o-mini: Advancing cost-efficient intelligence

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.782692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.058945Z digest=sha256:54c392efc907b7166e7e5daa91fbdc62129d07b03cad9c13fcb18666fb072e9c

Observation 377960dc-cb41-4e38-a7f4-d2932749b680 · outbound

This paper cites Sparse autoencoders learn monosemantic features in vision-language models.

Multimodal Model Diffing for Feature Discovery and Control Sparse autoencoders learn monosemantic features in vision-language models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.063731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.063731Z digest=sha256:860241b7fa8a0f9e9a3c9e86da8ce804034bde3801d7fd55d4365b5eca080f7e

Observation 1c619101-97ef-42f1-aae8-e70c8d36057a · outbound

This paper cites Towards vision-language mechanistic interpretability: A causal tracing tool for blip.

Multimodal Model Diffing for Feature Discovery and Control Towards vision-language mechanistic interpretability: A causal tracing tool for blip

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.766065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.068462Z digest=sha256:630a7800b34b7ad313b70eadcba68f7d27112081b03d7cec4694c76fac825318

Observation 94b2dead-18bc-4a07-af33-a10015ae3873 · outbound

This paper cites Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal.

Multimodal Model Diffing for Feature Discovery and Control Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.073278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.073278Z digest=sha256:d78de065b89781d0a8acd48911411e1386f73f32aae3450da63db83758c5bb99

Observation eacb6670-9287-4508-b7cc-10e65ab0b142 · outbound

This paper cites Visual adversarial examples jailbreak aligned large language models.

Multimodal Model Diffing for Feature Discovery and Control Visual adversarial examples jailbreak aligned large language models

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.750674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.078497Z digest=sha256:3fe5a0c6f4659e938b1f7278dab1600dcd573de83c487e4c58fde838bd4b4776

Observation d5566a4c-bbc4-4029-82c9-0268f9e08351 · outbound

This paper cites Qwen-Scope: An open sparse autoencoder suite for the Qwen model family.

Multimodal Model Diffing for Feature Discovery and Control Qwen-Scope: An open sparse autoencoder suite for the Qwen model family

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.733675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.083263Z digest=sha256:fadb60360f24157ae0077476cdd2200239c899c3505e93d751110a6ef489ae49

Observation fe447d69-874d-4cbd-b5a0-ea082f32f82e · outbound

This paper cites Learning transferable visual models from natural language supervision.

Multimodal Model Diffing for Feature Discovery and Control Learning transferable visual models from natural language supervision

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.088107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.088107Z digest=sha256:cae4872348aa8c14ee58059665c25c6dfc851ed44cf960c2a5a8863fe5217f0e

Observation eadbd835-4e5f-45e3-8fea-de3c5e559122 · outbound

This paper cites Jumping Ahead: Improving Reconstruction Fidelity with JumpReLU Sparse Autoencoders.

Multimodal Model Diffing for Feature Discovery and Control Jumping Ahead: Improving Reconstruction Fidelity with JumpReLU Sparse Autoencoders

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.092721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.092721Z digest=sha256:11276d32c0e406433e01036d57cfacb129153c3605ed9d106016b65b7af7ab98

Observation 4641de98-7e7b-4fc5-a4ab-6302312a8412 · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

Multimodal Model Diffing for Feature Discovery and Control Steering Llama 2 via Contrastive Activation Addition

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.097463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.097463Z digest=sha256:2050894d9234c1b95b3574b450e020461cfee049b383ab32b9e5be3ee7dccde6

Observation 8c892395-c64d-4597-88bb-e01c606b2917 · outbound

This paper cites Multi- modal neurons in pretrained text-only transformers.

Multimodal Model Diffing for Feature Discovery and Control Multi- modal neurons in pretrained text-only transformers

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.705410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.102224Z digest=sha256:d5d66071038896eca2e2279331824bbbdaafac321510efa323d06852674ba76a

Observation f298f62e-2df4-4676-8705-f70f9d84fe17 · outbound

This paper cites SteerVLM: Robust model control through lightweight activation steering for vision language models.

Multimodal Model Diffing for Feature Discovery and Control SteerVLM: Robust model control through lightweight activation steering for vision language models

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.688536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.106645Z digest=sha256:1ff8a97480c190da103dbe3f6dfd96988ef2ec6b0923434bcf7ccebe22595678

Observation 830f7044-5311-48c7-b4a4-e492efe747ce · outbound

This paper cites LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models.

Multimodal Model Diffing for Feature Discovery and Control LVLM-Interpret: An Interpretability Tool for Large Vision-Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.111721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.111721Z digest=sha256:c0f5734d39c2102017b854207f6edac0a9769d1873cc1962ec99b20150a306b4

Observation ea2aabab-eee9-460d-ba2a-2b5538ff20e5 · outbound

This paper cites PaliGemma 2: A Family of Versatile VLMs for Transfer.

Multimodal Model Diffing for Feature Discovery and Control PaliGemma 2: A Family of Versatile VLMs for Transfer

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.116317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.116317Z digest=sha256:e7ecdd5e35b4c5af877eb1cf8131bd34dd51949d55f7cf53b27bbbcec5599466

Observation 652c6242-0c83-45b3-84c8-a17a7d903b88 · outbound

This paper cites Daniel Freeman, Theodore R.

Multimodal Model Diffing for Feature Discovery and Control Daniel Freeman, Theodore R

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.122564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.122564Z digest=sha256:5a7fbc57862ca0f167e9d108fe62d00dc9f81dbcafc4370b4977de20160cffbc

Observation ad666547-3921-415d-9f0d-0f41acd4ac5e · outbound

This paper cites Li, Arnab Sen Sharma, Aaron Mueller, Byron C.

Multimodal Model Diffing for Feature Discovery and Control Li, Arnab Sen Sharma, Aaron Mueller, Byron C

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.659875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.127927Z digest=sha256:738b2423f9565fa202b7e998bf742632fef1ab5871d8a24e18c2ba8ef28859fe

Observation d37558b3-dc39-493d-bf70-b84cb8eedef5 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

Multimodal Model Diffing for Feature Discovery and Control Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.132540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.132540Z digest=sha256:1fd2cd08a62572e4e1c42bb729ec86247c92f4eae0718193ceb555331bc535a7

Observation 4f30b446-4ab4-4a1a-a778-10e1ba21a3ed · outbound

This paper cites Steering Language Models With Activation Engineering.

Multimodal Model Diffing for Feature Discovery and Control Steering Language Models With Activation Engineering

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.137372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.137372Z digest=sha256:3ec21a117e1f437a949d9e02241e88ead5670fe431a72e599132bcddf843d0ab

Observation a5e3ee40-f7fd-46f4-9da1-684f3115af2c · outbound

This paper cites Too late to recall: The two-hop problem in multimodal knowledge retrieval.

Multimodal Model Diffing for Feature Discovery and Control Too late to recall: The two-hop problem in multimodal knowledge retrieval

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.628662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.142365Z digest=sha256:63e7e30e44f30db06b9ed47a06453052bc925c721a4aaa31a3dd65d7aba4c254

Observation e97f989c-c33a-4426-8979-5210adbd71a5 · outbound

This paper cites How Visual Representations Map to Language Feature Space in Multimodal LLMs.

Multimodal Model Diffing for Feature Discovery and Control How Visual Representations Map to Language Feature Space in Multimodal LLMs

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.147155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.147155Z digest=sha256:44714c2602c9f11a4bec07448f7611c4a138405e999a617ad75d9b633875825f

Observation 9ddfa148-7c15-4249-9c65-72225f1a76bd · outbound

This paper cites Steering away from harm: An adaptive approach to defending vision language model against jailbreaks.

Multimodal Model Diffing for Feature Discovery and Control Steering away from harm: An adaptive approach to defending vision language model against jailbreaks

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.611779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.152172Z digest=sha256:6d2204f2a74a17976304d077ec9685b25a366558055af16f735779b083619b32

Observation 3ad407f8-f160-46c8-9936-e40a5904103c · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

Multimodal Model Diffing for Feature Discovery and Control InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.156829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.156829Z digest=sha256:355560a28679c80634f723a017df2d607f3230cc2010465cbda3c21937999194

Observation 3b14a2e9-6053-4aa2-9b6a-279769fe4d46 · outbound

This paper cites AdaShield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting.

Multimodal Model Diffing for Feature Discovery and Control AdaShield: Safeguarding multimodal large language models from structure-based attack via adaptive shield prompting

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.595098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.162853Z digest=sha256:3bb252c009b1a3c7024d2b6e94e9a0adcab75ce4c8a6c1e0f5783f2085a9da1a

Observation 47a8435a-2c41-4f4a-81e8-3c19919d4f69 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Multimodal Model Diffing for Feature Discovery and Control LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.167527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.167527Z digest=sha256:e160524e1f94664a8add971b91eff54345ab3fd89d55b9d6befa0261708d4b19

Observation 95f2b71b-43ef-4498-9931-0584ebf2d855 · outbound

This paper cites Qwen3 Technical Report.

Multimodal Model Diffing for Feature Discovery and Control Qwen3 Technical Report

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.172320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.172320Z digest=sha256:08bed0dddd98b680e6e94eec6108e66f68174d1dd0098afda863c24d429b9d05

Observation 581e9dcf-2c1f-4833-b068-0610fedbe8bf · outbound

This paper cites SafeSteer: Adaptive subspace steering for efficient jailbreak defense in vision-language models.arXiv preprint arXiv:2509.21400, 2025.

Multimodal Model Diffing for Feature Discovery and Control SafeSteer: Adaptive subspace steering for efficient jailbreak defense in vision-language models.arXiv preprint arXiv:2509.21400, 2025

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.177100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.177100Z digest=sha256:5664c2f7027b5ec54067ff31945646a824111ea191e854e794fea2b4319d1662

Observation c093c9d3-d5cc-4865-8131-636221f4691f · outbound

This paper cites Sigmoid loss for language image pre-training.

Multimodal Model Diffing for Feature Discovery and Control Sigmoid loss for language image pre-training

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.181619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.181619Z digest=sha256:5a6ab7bf3ca286489fb211393525b46de2e4230c14e30a80dbae03654ea04615

Observation 112922d6-827d-42aa-a3d2-3aea70380e8f · outbound

This paper cites Towards Best Practices of Activation Patching in Language Models: Metrics and Methods.

Multimodal Model Diffing for Feature Discovery and Control Towards Best Practices of Activation Patching in Language Models: Metrics and Methods

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:56.186065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:56.186065Z digest=sha256:56806c9dd7074fbfa4cd4c0d991650d6773b581c8eaab3fe3fd41f630492eda1

Observation f2b68228-1a9b-4451-b843-561f9a915118 · outbound

This paper cites Cross-modal information flow in multimodal large language models.

Multimodal Model Diffing for Feature Discovery and Control Cross-modal information flow in multimodal large language models

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.566959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.190937Z digest=sha256:850440320b291a8e76f0cce330dc46d69533a17b0574bcf466f099170763b186

Observation 2619e432-0229-4ce4-9384-2c2445bcae20 · outbound

This paper cites Multimodal situational safety.

Multimodal Model Diffing for Feature Discovery and Control Multimodal situational safety

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.549778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.195809Z digest=sha256:c057e5cffaa0b96526d96ed7692271618c198f027c232724219cbe432b3980f2

Observation e7d95dd4-96b3-42a8-bfa2-b16e8a5e4b3e · outbound

This paper cites Relocated.

Multimodal Model Diffing for Feature Discovery and Control Relocated

Reference 95

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T04:17:57.530047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.201760Z digest=sha256:0a4ba31574395d2ecbd3e4885deeeae4feeaf886ca0b607613564574c2608984

Observation dd40ce95-4dfb-4b0e-bb9f-d4bfd52a465d · outbound

This paper cites an unresolved cited work.

Multimodal Model Diffing for Feature Discovery and Control Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:17:57.512239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.208372Z digest=sha256:fc92f09c4bf714ac97fbcda110811ce11524977a3eda87f346109a35e6b3e919

Observation 7bbbfb87-1a52-4a94-8530-dc262d50d734 · outbound

This paper cites an unresolved cited work.

Multimodal Model Diffing for Feature Discovery and Control Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:17:57.494613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.214057Z digest=sha256:241b679e4dd52073264f87dbf8b7c21b9d1487cff13cec1c7c8dd176dda81863

Observation 418b9983-293a-4b05-83aa-d201d05c83b7 · outbound

This paper cites an unresolved cited work.

Multimodal Model Diffing for Feature Discovery and Control Unresolved cited work

Reference 98

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:17:57.478120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.219618Z digest=sha256:3d88ea8db3d332ddd442ac82c0ad79ee585892db8f343792f8f294a8327ce1be

Observation 4447cbf9-6ef0-4020-9860-8acf0b877876 · outbound

This paper cites this neuron activates for.

Multimodal Model Diffing for Feature Discovery and Control this neuron activates for

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:57.462367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-11T04:17:56.224644Z digest=sha256:ecd41ba3f43e7a4e8095d75bb892658bc7ac8fe315d7a33fbe9b329520831c31

Pith citing papers

No inbound Pith citation observations are available.