Pith. sign in

Paper Citation Record · LEDGER

Robust Multimodal Large Language Models Against Modality Conflict

As of 8 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 2 inbound Pith citation observations for arXiv:2507.07151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07151 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:02:35.018852Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T01:00:28.684512Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:18:56.608098Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6a81bd36-c9fc-4228-8559-153a6cf28398 · outbound

This paper cites Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs.

Robust Multimodal Large Language Models Against Modality Conflict Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.987671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.987671Z digest=sha256:c1907608fe632a707e1f35ed21eeb5de6851b16bbb969a786e79d8f099421b39

Observation c8fde335-0da2-4685-be08-d6a92b196ec7 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Robust Multimodal Large Language Models Against Modality Conflict Proximal Policy Optimization Algorithms

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.993758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.993758Z digest=sha256:8532f45d00efaa4aacf1543a500d73e22ca2998f93cdff16cfc9a79da009f693

Observation 6c360178-5c32-449c-9339-58db71ec1181 · outbound

This paper cites Knowledge conflicts for LLMs: A survey.

Robust Multimodal Large Language Models Against Modality Conflict Knowledge conflicts for LLMs: A survey

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:02:35.375264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:02:35.009128Z digest=sha256:4686c0631837c0bd0d6dd57ffaefff219483bdf4d7cd5e12562d6a992f5b8ddc

Observation 824ec269-4595-4289-b25a-17bc9e75a33b · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

Robust Multimodal Large Language Models Against Modality Conflict Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:35.013941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:35.013941Z digest=sha256:8b172c4c70aab769c61d946b203b0ed2fbcb483eadeecdce29c30eb08aa96343

Observation 63c38f93-dfe9-4510-9ac0-fee91f0bec99 · outbound

This paper cites Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models.

Robust Multimodal Large Language Models Against Modality Conflict Unraveling Cross-Modality Knowledge Conflicts in Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:35.018852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:35.018852Z digest=sha256:8407998851e1149de51b994cd6a16437f4e895ae1de7cd0ff3f6fda278666df4

Observation 42c00eac-12e9-4bea-876e-32f40293ee15 · outbound

This paper cites Large vision-language model alignment and mis- alignment: A survey through the lens of explainability.

Robust Multimodal Large Language Models Against Modality Conflict Large vision-language model alignment and mis- alignment: A survey through the lens of explainability

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.998854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.998854Z digest=sha256:c0f617fc22494c60151f0bc97b60570ed0284bf507bd2c47bc718adc3aef916e

Observation 66490eca-cb45-49fa-9753-2f6979fbfde8 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Robust Multimodal Large Language Models Against Modality Conflict Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:35.003935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:35.003935Z digest=sha256:d53bc22f1a0b76c6bfc4aa8d3d2dee593d2c65199e3915b5e972d3b1b0cfa646

Observation eef14c0d-2e58-4ecd-a39a-9547bd6f71b4 · outbound

This paper cites REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization.

Robust Multimodal Large Language Models Against Modality Conflict REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.976580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.976580Z digest=sha256:78d5916f08186f7bda21dc73ef4d548e74462c233f0745a8b5f9b4aa80d936ee

Observation ea0551ff-bf15-4c4d-b400-1820d21f587d · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Robust Multimodal Large Language Models Against Modality Conflict Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.960675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.960675Z digest=sha256:b52db73ec61ffcc8d77a85ed472b6a01b80d788f15f9405aad4acc559e902429

Observation e4cf7d51-617f-4236-bf7c-c7701e49d89a · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Robust Multimodal Large Language Models Against Modality Conflict MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.965966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.965966Z digest=sha256:9340a0db1a260b33d8399925311d87499dc5a6774dd755ae52e9ad080a159420

Observation a6a9822f-a255-409c-aa8e-385e8ac1b09d · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Robust Multimodal Large Language Models Against Modality Conflict LoRA: Low-Rank Adaptation of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.971396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.971396Z digest=sha256:03cb5da48207455d8a87b5df02f71ab9fd9b9d583403d98940b0544058000eda

Observation 42e835f8-6907-4cc8-ae47-3db44957f3d4 · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

Robust Multimodal Large Language Models Against Modality Conflict OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T19:02:34.981872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:02:34.981872Z digest=sha256:1f18462211789e2ee02b780bf6db88aa7b4c805fe8ba9fed4a9af7cc67500cd7

Pith citing papers

Observation 3268c7ed-89ce-4f77-801e-bbd66db5bb4c · inbound

Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability cites this paper.

Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability Robust Multimodal Large Language Models Against Modality Conflict

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T01:00:28.684512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:00:28.684512Z digest=sha256:ed4e3479b5bcef0b517c98ad223616d920b0447a6f78ced806eaceee329ebe34

Observation 82844a20-d46b-4cca-b895-8d54fd7ae614 · inbound

MLLMs Get It Right, Then Get It Wrong: Tracing and Correcting Late-Layer Textual Bias cites this paper.

MLLMs Get It Right, Then Get It Wrong: Tracing and Correcting Late-Layer Textual Bias Robust Multimodal Large Language Models Against Modality Conflict

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:56.610055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T01:28:50.021432Z digest=sha256:c4a2071ce1c50a970b8aad3c0c6c8673ef77634e85fff917e67cabac6f40b762