Pith. sign in

Paper Citation Record · LEDGER

mDPO: Conditional Preference Optimization for Multimodal Large Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2406.11839.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.11839 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:22:59.547939Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:40:00.900784Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5697439a-5277-408c-8e38-ce672ccb5e2e · inbound

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning cites this paper.

ASPO: Adaptive Sentence-Level Preference Optimization for Fine-Grained Multimodal Reasoning mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:59.547939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:22:59.547939Z digest=sha256:09bb46faedfccf48d879835a8524501e4f55cf24abc9a40c757743c27a1b6302

Observation abfa1f5f-c24f-4664-a279-ea700245fe5c · inbound

LPOI: Listwise Preference Optimization for Vision Language Models cites this paper.

LPOI: Listwise Preference Optimization for Vision Language Models mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:45:57.282620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:45:57.282620Z digest=sha256:a1813b323e801b52332cf5bd9511bd200d24fe5f08a1663721837d7c4caa874c

Observation a1fbe716-a90d-43ec-913a-c1820312abbb · inbound

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning cites this paper.

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:19:32.483603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:19:32.483603Z digest=sha256:863ef5e8338ee62773f6f60c360646451a0f1176df2f7a8f8e01f732b28dd078

Observation 51a431a8-93ad-4dd7-b171-f5796f46097a · inbound

ReFoCUS: Reinforcement-guided Frame Optimization for Contextual Understanding cites this paper.

ReFoCUS: Reinforcement-guided Frame Optimization for Contextual Understanding mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:51:29.258175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:51:29.258175Z digest=sha256:8afab3e66f332d8bb4b6e1f50af62acf2b4e4501e008a781f1319b03b8bd0693

Observation 52fc45d8-d8ae-47c0-8146-e8dab32c6959 · inbound

DPO Learning with LLMs-Judge Signal for Computer Use Agents cites this paper.

DPO Learning with LLMs-Judge Signal for Computer Use Agents mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:11:47.992971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:11:47.992971Z digest=sha256:3770a1fc22575add86539574d9b1bffbb2fdb891048bb681d4a06870b9483fb9

Observation 88349501-5533-413c-8a10-33c1c19c004b · inbound

Explicit Preference Optimization: No Need for an Implicit Reward Model cites this paper.

Explicit Preference Optimization: No Need for an Implicit Reward Model mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:18.024969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:40:18.024969Z digest=sha256:98306c7728906f9a1ed07b1e5a62a5207ec63ffd2882e931d8aceb0e06290914

Observation a8db0e4d-af71-4ae6-8e6c-67244ac9b2e8 · inbound

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models cites this paper.

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:09.378681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:09.378681Z digest=sha256:16d5080c68cf30db875e15bf54ed9e9f40d7eeccaa7b352667ebd8bb24012e02

Observation a82b1c07-6d07-4831-ad0d-e46326757cc6 · inbound

Controlling Multimodal LLMs via Reward-guided Decoding cites this paper.

Controlling Multimodal LLMs via Reward-guided Decoding mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T19:52:50.338843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:52:50.338843Z digest=sha256:20b694b82360f0d431471b456441bbf5fcb3aebdd3ceaac4acb1adcca9331953

Observation 152d178c-ba8d-4154-b005-e55fe4622cc7 · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 193

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T23:04:01.769524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:cd00f682cbbc6f7978bfbfd1df376d3c88ee5f49cc9fddd88d78da7dae2a8c30

Observation 219c162c-0b6c-4658-885a-4463d8765090 · inbound

P$^2$-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization cites this paper.

P$^2$-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 104

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:06:27.158617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T11:13:37.291834Z digest=sha256:0d99920d7e8ec2b01f5c859b46e5d6a2d8e5a81494e920a44914df942e25c05a

Observation d8011104-e685-4cfb-9320-b8bf88c1d33d · inbound

Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination Mitigation cites this paper.

Steer Where It Matters: Token-Level Visual-Sensitivity Steering for LVLMs Hallucination Mitigation mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:06:29.645786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:22:40.055153Z digest=sha256:da7bffd4a54dc94a37006ab441c872600662f825b084b128598d456b45863637

Observation 98ac3165-8345-4bc2-bd97-17ee65456232 · inbound

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning cites this paper.

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.902288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T23:29:24.520537Z digest=sha256:ac8fb09c5b32edffbc33e10e5463c5c8cde53ed594fdfe080f72dd120aa6b8bb

Observation 250c4487-1964-483a-ad59-36bd8cad9b5c · inbound

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs cites this paper.

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 68

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.328171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:35:08.200219Z digest=sha256:db8edf837104e4a2dbeae1300e27487447a28c921154c9be402e250e3c48259a

Observation e2c81915-7159-47d7-99cd-15a253b2b8ed · inbound

Adaptive Perturbation Selection for Contrastive Audio Decoding cites this paper.

Adaptive Perturbation Selection for Contrastive Audio Decoding mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:07:12.153703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-02T16:59:50.615875Z digest=sha256:307655d196fd6eed9b9697c2c274e41bc5f0dacd13ad5196173cc5b0fd7108b9

Observation 54b4d8b3-3585-4864-9bb9-2bdb789a45d3 · inbound

Adaptive Perturbation Selection for Contrastive Audio Decoding cites this paper.

Adaptive Perturbation Selection for Contrastive Audio Decoding mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T04:34:52.059708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:34:52.059708Z digest=sha256:c4d279cd8d5bd7d3aa0f49165f33b9d9f23d74972d0d9e764c90670b51f0798c

Observation 2d1d02aa-cf1e-4d2b-89a6-9aaecdf85394 · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 247

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:7f0ec3b1e4c5dfe8d939519c40b23bb3cba550fc0ba6abcacd1987b2548a2a56

Observation 6fe0bd4d-c33b-4840-8544-a6898218488f · inbound

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs cites this paper.

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T04:15:38.709576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:15:38.709576Z digest=sha256:84fbff6992b3628f992d65fccdfecfb8c6c2f90494b7da24f36e6e9be97e604b