Pith. sign in

Paper Citation Record · LEDGER

MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2505.24871.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24871 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T08:15:52.772665Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T22:24:00.434533Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bb4826de-a215-4f05-b0f8-2f7d09409a9e · inbound

Perception-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Perception-Aware Policy Optimization for Multimodal Reasoning MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:04.799535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-19T05:11:54.685897Z digest=sha256:be16afcd43b538d9db17d16ad08d41b65f61d0e48142c3dfd64769ac205f948e

Observation 2917cdcd-d538-4c7a-851d-80e5c0bc0eed · inbound

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models cites this paper.

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T08:15:52.772665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:15:52.772665Z digest=sha256:a2e743fc5aa03e9939fc96c37051cbb6b01816b26787e7bc0fa026257e15e5e9

Observation ed54ce13-893a-4ea7-b882-229b449251d0 · inbound

Multi-Task GRPO: Reliable LLM Reasoning Across Tasks cites this paper.

Multi-Task GRPO: Reliable LLM Reasoning Across Tasks MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T04:17:27.630062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:17:27.630062Z digest=sha256:ecf4ed32af48b9397a1343bc4e8e123d62eb93a822730467cc3944a40a9af05e

Observation 2305ebed-33fd-4d4f-bbe1-e088f0d8af4c · inbound

Harmony in Diversity: Multi-domain Contrastive Policy Optimization for Large Reasoning Models cites this paper.

Harmony in Diversity: Multi-domain Contrastive Policy Optimization for Large Reasoning Models MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:24:00.436153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T22:18:10.508034Z digest=sha256:ecf87172bdb76ccff81a2c127092d7db62dc1894253d25c6985b0a42c853be4a

Observation 5caa1d1d-90d8-4143-9426-0394621fd17e · inbound

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning cites this paper.

Trace: A Taxonomy-Guided Environment for Multidomain Visual Reasoning MoDoMoDo: Multi-Domain Data Mixtures for Multimodal LLM Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:47:24.668308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:47:24.668308Z digest=sha256:0d08b547bf25ecd81c67f4ac5e78c835b031239d0207501e7456c0baeea876ed