Pith. sign in

Paper Citation Record · LEDGER

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

As of 15 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 4 inbound Pith citation observations for arXiv:2509.22415.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.22415 v4

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T14:58:41.930696Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T13:51:30.008232Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:17:26.204542Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation edd6808e-8f03-42ae-bb2d-ad4d16c8f989 · outbound

This paper cites Quantifying attention flow in transformers.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Quantifying attention flow in transformers

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:36.640029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:36.640029Z digest=sha256:2724d719557bdee8f12427baa11837aeb1969ffa733b3b6075292b89af0f4c0b

Observation abfb8110-11d4-4276-bc08-6fd88d795953 · outbound

This paper cites Attnlrp: Attention-aware layer-wise relevance propagation for transformers.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Attnlrp: Attention-aware layer-wise relevance propagation for transformers

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:36.714273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:36.714273Z digest=sha256:89533e5aed0ad2a8425ae8c1a7fce4541bae805b15b866d62842ece4aadcabed

Observation bf34d5b2-f9b4-42e3-b99c-e2632bd4d4b8 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Flamingo: a visual language model for few-shot learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:36.864905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:36.864905Z digest=sha256:b56c64eeaeb2ac889798a08e4e89083cddd06220a1895d93adde4e428ac9ca2b

Observation 152e38f3-4ebf-4464-9efb-ac784ddef259 · outbound

This paper cites Xai for transformers: Better explanations through conservative propagation.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Xai for transformers: Better explanations through conservative propagation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:36.984919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:36.984919Z digest=sha256:fee3a440d27c10fbf4de3b8eac611faab18a5d26a6fc010e7222e5b3230094f8

Observation c2485d37-b6f7-4ad2-b452-8ba701e091dc · outbound

This paper cites Vqa: Visual question answering.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Vqa: Visual question answering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.085767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.085767Z digest=sha256:e46be2f052eb689bed20fbf8061ed9dd2fabcfbf2434cf3ead39e0b490fdaeac

Observation 7f217119-0866-4247-8a30-cbc057680002 · outbound

This paper cites Lvlm-intrepret: An interpretability tool for large vision-language models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Lvlm-intrepret: An interpretability tool for large vision-language models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.164745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.164745Z digest=sha256:074ce11d5fadc2e7924136a153e8eb8492cb4a05a3572173b2c8b30feed3085e

Observation d7655173-a537-48d3-9a1e-5fea2d13798b · outbound

This paper cites Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.305848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.305848Z digest=sha256:e960f9ead0de6e6c763387d65acdddd4753e0ebff8d29e00d323447e49dbc5d7

Observation cfdadec1-1bae-4b93-9db0-91c600144b44 · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.449797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.449797Z digest=sha256:9aa9a6838d1b12ca3c326d7871c7a5fd13c889629bee4dda5680f7d9bbeaaab6

Observation 01efeb9c-10ad-498d-8f54-d83ab0cefec8 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.574822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.574822Z digest=sha256:ea4dd6322184ce8402525693d65969bfc98eacb731e121054ba8cebf024643d3

Observation f2811cf3-8a57-4d68-a82a-fc95994ed473 · outbound

This paper cites A comprehensive survey of deep learning for image captioning.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models A comprehensive survey of deep learning for image captioning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.784819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.784819Z digest=sha256:5df9aa547ba080cee59af202ae24cd51844b800ba8a9aaf6129b737e2ba7ba68

Observation ba9802a2-c408-4e9b-9f1a-1e1e868b4c31 · outbound

This paper cites Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.914746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.914746Z digest=sha256:dc79bb27ae3386232ce435b673556cf33eb49a828bbde6756f6426ffe51d5d4d

Observation bac4670a-8c94-4185-91c6-65176859b581 · outbound

This paper cites Layercam: Exploring hierarchical class activation maps for localization.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Layercam: Exploring hierarchical class activation maps for localization

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:38.126706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:38.126706Z digest=sha256:ab5508ed89b19f3d4618b9ead24e486d1451912226c7c5e236a6e400c4803720

Observation c4c87463-6172-4101-ba6b-fc14b2d04af0 · outbound

This paper cites a ldchen, Alexander Binder, Gr \'e goire Montavon, Wojciech Samek, and Klaus-Robert M \.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models a ldchen, Alexander Binder, Gr \'e goire Montavon, Wojciech Samek, and Klaus-Robert M \

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:38.253254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:38.253254Z digest=sha256:84f7f8db2361645292b8cfb8b28269a96f42350e2d180dbaf2b5c00ead0d0f1b

Observation ee607a6c-424c-4103-8a9b-d5ac4adba379 · outbound

This paper cites Token Activation Map to Visually Explain Multimodal LLMs.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Token Activation Map to Visually Explain Multimodal LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:38.554863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:38.554863Z digest=sha256:ca8e6cbf224f67b2f5eae3174f1917a1120b6f792b3074a307482611fcb5e076

Observation 42ed0f55-deae-4c2d-ad9e-6ee029756c0f · outbound

This paper cites A closer look at the explainability of contrastive language-image pre-training.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models A closer look at the explainability of contrastive language-image pre-training

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:38.704820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:38.704820Z digest=sha256:1c2e7f69256943e7c16349dc8d62c3114403c26b700ac989f5447f154ac40c4b

Observation 45708298-6f54-4636-bef6-1ec0ff012a26 · outbound

This paper cites Imitated detectors: Stealing knowledge of black-box object detectors.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Imitated detectors: Stealing knowledge of black-box object detectors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:38.834822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:38.834822Z digest=sha256:50543b37544976e6a322db40dac6bdc2f9a301331567954bf422fad81cb3f976

Observation c0aa5596-2fd7-4009-8a55-59bc2d54052f · outbound

This paper cites BadCLIP: Dual-Embedding Guided Backdoor Attack on Multimodal Contrastive Learning.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models BadCLIP: Dual-Embedding Guided Backdoor Attack on Multimodal Contrastive Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:38.944873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:38.944873Z digest=sha256:5e9aab16058061b8cef209b65066d2d4a849e1532341fbb9db48f562c8515dfd

Observation 4cedad88-ca80-4441-81fb-87b8e8d359a2 · outbound

This paper cites Revisiting backdoor attacks against large vision-language models from domain shift.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Revisiting backdoor attacks against large vision-language models from domain shift

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.095066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.095066Z digest=sha256:717c06b5efa15b8dfe46c41555991cd36ba2f888d48c919136bf0be9bb7cbb7f

Observation 359f2b2a-fd58-49fc-bdf2-aa81da3cc33d · outbound

This paper cites T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models T2VShield: Model-Agnostic Jailbreak Defense for Text-to-Video Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.157286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.157286Z digest=sha256:367ed058d5f97ba354e3efa41218fd94203d7ee86bb8849734b6cdc47b8cc124

Observation a9e50e97-3454-424f-80e5-4548b1b2298d · outbound

This paper cites Microsoft coco: Common objects in context.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Microsoft coco: Common objects in context

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.244816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.244816Z digest=sha256:72fc02b5ed8f752dd4003b9e73e21d9acd951bdf7eda3df4f9282fc222f843d4

Observation d4a11e8d-3a6c-4e97-bb1c-91bcaf86be32 · outbound

This paper cites Agentsafe: Benchmarking the safety of embodied agents on hazardous instructions.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Agentsafe: Benchmarking the safety of embodied agents on hazardous instructions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.514777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.514777Z digest=sha256:17392e01b2d402d1793879f5440bd764751ba6319570bc13a024d05649f82e30

Observation 420c2142-2f54-48d9-9f5f-0f11eea76121 · outbound

This paper cites Visual instruction tuning.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Visual instruction tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.632162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.632162Z digest=sha256:ef410036b3b8ea31f090568653a5b561244bba6484ca5f35f4e964863fc790af

Observation f486202e-5d79-49fa-ab5e-508403e590c0 · outbound

This paper cites Introducing mpt-7b: A new standard for open-source, commercially usable llms, 2023.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Introducing mpt-7b: A new standard for open-source, commercially usable llms, 2023

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.774940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.774940Z digest=sha256:2711cdee0242a46fa57819849ea72479924943af19df7a2fcd9dab752e235b36

Observation bd5ab252-c52e-4191-a1e6-36781453d342 · outbound

This paper cites Interpreting GPT : The logit lens.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Interpreting GPT : The logit lens

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:39.894820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:39.894820Z digest=sha256:f4bd582ff2569da2d396c7d3db09989b9c585aa5467394907de65a81f2e449e6

Observation 6ab32022-082f-4766-b351-7c12071380f5 · outbound

This paper cites Glamm: Pixel grounding large multimodal model.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Glamm: Pixel grounding large multimodal model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.044812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.044812Z digest=sha256:1d075f6b64d515f189804481f5ff264d17de3159fa55df01bff943904fd0ad77

Observation 368a8580-d972-4a70-87af-c473db46328f · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.135308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.135308Z digest=sha256:8301b61ec36a301054167f869029cf86dacf601e42a66a15511a749d85dbc3d7

Observation ceedae73-1b4e-40e9-a871-c839a53dec35 · outbound

This paper cites Releasing 3b and 7b redpajama-incite family of models including base, instruction-tuned & chat models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Releasing 3b and 7b redpajama-incite family of models including base, instruction-tuned & chat models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.248234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.248234Z digest=sha256:8ae1e9b118dc9a4404bd2ef274b4cf7e8ab4807d8c517be3427a4722a38f9912

Observation cd6462a4-3d92-42cf-968f-a224d15b9ba4 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.334874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.334874Z digest=sha256:06fb45de861a1df3e2cc03cf44479a0da374b209438963f36d6376366e1faf1e

Observation b2b598e6-7327-4aec-9ee9-1d55d4d98312 · outbound

This paper cites A similarity measure for indefinite rankings.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models A similarity measure for indefinite rankings

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.406592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.406592Z digest=sha256:4274936d86c410d2313f38e85e293712b45abec446896c336b3e85dabc826d6a

Observation 8ac6b1a9-183a-4331-a43f-a5ca14aefe4a · outbound

This paper cites GenderBias-\emph{VL}: Benchmarking Gender Bias in Vision Language Models via Counterfactual Probing.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models GenderBias-\emph{VL}: Benchmarking Gender Bias in Vision Language Models via Counterfactual Probing

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.532601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.532601Z digest=sha256:c0010d4f9e9a110a121f4b2bfddefbdbfca525be01975bc05efe3ded7cc686d5

Observation ded2961a-f7f7-4119-b76f-1ff743a4f816 · outbound

This paper cites Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.684840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.684840Z digest=sha256:af6f08baeda14e7b42253dfe3ba7a5831ee106a1f77956a343ed4ad7be518d26

Observation d5b73532-6df5-4685-854b-badf749082d0 · outbound

This paper cites SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.772114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.772114Z digest=sha256:e341aae9989de87a3b01f784ec86390f9ccd5452abc3c000fb03a4fe3d452d90

Observation e4b34017-fd2e-46ab-84b7-de38e5014b22 · outbound

This paper cites Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Jailbreak Vision Language Models via Bi-Modal Adversarial Prompt

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.894740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.894740Z digest=sha256:6e9141fc012cbb2010175352e31e0057c9217479e55ef6a5f090e8dd964b45c6

Observation d959c428-ffb2-4b39-a8e2-ed0d461d1b1f · outbound

This paper cites Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:40.973955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:40.973955Z digest=sha256:010750001bfd1c8108bf3659166f90ef2bb3385ef2dd999498701edff26d6f1d

Observation 0412c0fb-1f5f-4c58-a4d5-59aa5729db60 · outbound

This paper cites From redundancy to relevance: Enhancing explainability in multimodal large language models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models From redundancy to relevance: Enhancing explainability in multimodal large language models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:41.085453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.085453Z digest=sha256:df42cf0d9695d45efc6feeb052ee51ff90c77e612a4b2e804157091c31c9c247

Observation b10d9ec3-a928-4483-9b21-64edc7ea1b33 · outbound

This paper cites Learning deep features for discriminative localization.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Learning deep features for discriminative localization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:41.170192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.170192Z digest=sha256:421b29ddbe0f0f44e7f5d5527de985b9c8b89ae137e47fc760fc6fb39eb0cdb8

Observation c66b70d3-e0d2-417b-8d63-80fab847d6ae · outbound

This paper cites Openpsg: Open-set panoptic scene graph generation via large multimodal models.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Openpsg: Open-set panoptic scene graph generation via large multimodal models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:41.290378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.290378Z digest=sha256:cbab16e029d7e0c53f2aaab8058c0e436310104467e4fb8b5066f5c7ff4c5d7c

Observation d7a4f2f8-c28f-4a83-9b35-8a9e401b1bc9 · outbound

This paper cites write newline.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models write newline

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:41.464746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.464746Z digest=sha256:ff91ced3d99d5720d5d8cbb87d4b0ed6e694860c7ae641d9d822e8717693c80d

Observation b0a9a974-ade6-4253-bb97-88e46c13800d · outbound

This paper cites @esa (Ref.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models @esa (Ref

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:41.727878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.727878Z digest=sha256:76c067f80f2514cca2a7bc8d9d928f900d9b88d173f73c6116f2eb9dc66b84db

Observation 70b12ab8-4a0d-4e4b-8ffa-f9c2195fede1 · outbound

This paper cites an unresolved cited work.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:41.864889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.864889Z digest=sha256:c1f6969aece30971e0e077b53c5f3bd1984a1c1195dd499fa6fc707d9876b49f

Observation 0a43367e-89d0-4626-b19a-29e5544aabe5 · outbound

This paper cites traffic" and.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models traffic" and

Reference 41

Resolution
malformed identifier
no resolver link, observed 2026-08-04T14:58:41.930696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:41.930696Z digest=sha256:35433a7ad84ed1f956f849fb540e656e3d951e1bf5dca12c1de2132c24b9d6bd

Pith citing papers

Observation 3ba56f85-3d55-4672-bfb4-ae7e7e20a6c9 · inbound

What if? Emulative Simulation with World Models for Situated Reasoning cites this paper.

What if? Emulative Simulation with World Models for Situated Reasoning Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-15T13:51:30.008232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:51:30.008232Z digest=sha256:a3424e65d055e133b0215b304bb65056dc59a20f71a99b20de5f1e20c04bae52

Observation 38d1b9c1-92c1-4c09-9345-f8178441dbba · inbound

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference cites this paper.

MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-15T02:21:03.503340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-11T01:29:26.298131Z digest=sha256:8aed8e35df17467754f20ecf37a8579d148cbced6eb7df5599a5a95f6a7146a2

Observation d5e45228-c43b-4eda-941e-6e2dfb7ef000 · inbound

Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads cites this paper.

Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-15T02:21:03.503340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-28T01:54:13.927736Z digest=sha256:92e1f7673843ab8f78d9a70014a9eba19466a9e2fa7a9275d5109da7a333fda4

Observation db50e166-6b9e-46ab-ba4c-0ff270c67d5c · inbound

When Correct Decisions Hide Internal Stress: Decision-State Probing in Multimodal Language Models cites this paper.

When Correct Decisions Hide Internal Stress: Decision-State Probing in Multimodal Language Models Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-07-15T02:21:03.503340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T19:02:12.613002Z digest=sha256:ef054fefb2b49af04d537ca22ef9a30e141065af2cc7982a01800924709437d5