Pith. sign in

Paper Citation Record · LEDGER

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation

As of 7 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2506.17664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17664 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:34:56.649817Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-11T01:49:15.136031Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T16:01:22.762740Z

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a01c9e40-12ba-4f38-aed7-c5675210c69f · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.081734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.081734Z digest=sha256:b5cbbeaade57030955d6bc65c16e1e52e590d4ae8f1754f99c76af3371f22a2f

Observation 5381a6ff-955e-46b2-a247-d81895bd4159 · outbound

This paper cites Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.164743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.164743Z digest=sha256:aa3341ac0c0264d48caa04b41fde0e485c1ff4bd964b6340b0ae7d0a64245239

Observation 3acb9a3f-0ac5-4fc1-9016-dfa90bb4677d · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.278265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.278265Z digest=sha256:3f4bb2d20ac015eabb9f06527d15c37bb33bb5a2053c2cf4f55fd971389a1b15

Observation 8b2b0a8f-af0a-4306-ae4a-dd4706c9bad8 · outbound

This paper cites GPT-4o System Card.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation GPT-4o System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.465653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.465653Z digest=sha256:aec7c7ee6af99002e42540506868e753e7597b36bab71051d9fa80ccc28823f8

Observation 58ef4792-b290-4307-9c93-f8f8618fade0 · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.674752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.674752Z digest=sha256:00ca1d80727d7a134477bec30550faac6b75f9b7002e8c5eb903ae421a225b66

Observation f5974394-d72d-41d7-8774-134f78e52d20 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.735935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.735935Z digest=sha256:02e971ad3f3cc8fcf62a9f8b7abd394987cfc50c004ee526c6f6cb9150ecb8b1

Observation 6876b6ee-b62b-4f54-bc7e-060ab3694c6a · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.885669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.885669Z digest=sha256:2e762cc7e72162c71692afe04d637a50e4d06963462f4fc299590082b8a5b01f

Observation b08c0987-0469-4868-ad41-5cf26c981d82 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Gemini: A Family of Highly Capable Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.934743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.934743Z digest=sha256:050f2c400cdf1c9ab50663262d3eb2a6aa3ec06f8e80e20ad14062a504f8ac7f

Observation b6bd49ea-46fc-4c65-836e-cf82d369b442 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.992900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.992900Z digest=sha256:24766cb85eaa1014a8ff8c0f15e22e7c6cfeece71c7b1c398553c11282b41f34

Observation ded3c166-299f-491a-8d8a-cf7f8be290b8 · outbound

This paper cites Caption Anything: Interactive Image Description with Diverse Multimodal Controls.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Caption Anything: Interactive Image Description with Diverse Multimodal Controls

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.110735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.110735Z digest=sha256:b0785d8eb87c6a632a94073379ae9b6b10f47e4bc6c30293b9b731fce0bc407b

Observation 0bdd5568-3306-4e1c-9e9b-31e075c5403e · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.275773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.275773Z digest=sha256:3d16e184f5c14d8a6a2915f9e28252714aa6d884f39f8e5022c56915cecdebc6

Observation abe0cc92-da89-4367-a161-33631ac7882e · outbound

This paper cites mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.417716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.417716Z digest=sha256:ceac1c699d12bfbd86bd98169811008bad576741efc697555a3084b680b04302

Observation a95efcf4-84f7-4b69-8af0-693b4a982f1e · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.541827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.541827Z digest=sha256:ab1ee50a01d9fbc464cdf1fb6e5a9a7d40f48fde6bbae33127a6a7ab42926395

Observation 921ef3b2-bdf5-4337-85df-c2d8e04ff6f1 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.649817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.649817Z digest=sha256:28b6afbb758ece6f562528f68cec02be5791a8ca617c8e42c7a82f654ffe824a

Observation 39332d3d-a380-49ed-9974-bd42dc128f82 · outbound

This paper cites Object Hallucination in Image Captioning.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Object Hallucination in Image Captioning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.813461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.813461Z digest=sha256:01f39a27df9ac666cf39916a2956b1a962b8ae61ca69219d85be6cd6d573fc06

Observation 751df6f5-9b7c-486d-b97a-77f3358bb160 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.544749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.544749Z digest=sha256:e2ceee8558036d3a4d59d19c071f0532d42d915e0a0d332cfa37711b5e0703b4

Observation 400765dc-0cac-40bc-a2aa-1021cce2aab4 · outbound

This paper cites MultiModal-GPT: A Vision and Language Model for Dialogue with Humans.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation MultiModal-GPT: A Vision and Language Model for Dialogue with Humans

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.367724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.367724Z digest=sha256:531d3ec223b577c44c798e1f9b363b3dabd5c1baf7f9a207557120f7fe67c1eb

Observation 629c7d3e-52de-4b5f-9070-f98089e618c2 · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:54.937267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:54.937267Z digest=sha256:fe635cc38158674048a159614c6356a1dedbbe88f6394eec6e91782e1e5c5918

Pith citing papers

Observation a6efc5ab-82e8-46db-9ee7-616d25da5476 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:01:22.777348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T18:53:06.494640Z digest=sha256:5fdb811ee6b5bb9224b41cd837aaf6262c67a746208615b57d3e6d5ac2f53536

Observation c1436332-d86a-4610-9ff0-277a199108ba · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.532867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:49:15.136031Z digest=sha256:aea7f56ca2df39145a9593ac11a1f13fed8791ff6021a10034953c4433c4b532