Pith. sign in

Paper Citation Record · LEDGER

MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2407.12709.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.12709 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:32:11.897837Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T14:42:57.008976Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cbe97125-4acb-446e-9e35-6a0d1b8cdbac · inbound

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning cites this paper.

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:42:57.013953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T14:42:56.565621Z digest=sha256:0b3bf8c0213050772f050c0ebb752e0b6321076d4c6c0bddfe30105f23af38aa

Observation e61748c8-ef1e-46a9-b396-87264baf6307 · inbound

MINT: Multimodal Instruction Tuning with Multimodal Interaction Grouping cites this paper.

MINT: Multimodal Instruction Tuning with Multimodal Interaction Grouping MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T11:32:11.897837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:32:11.897837Z digest=sha256:c434533850ca9c89cd651ad8c3220794d7abbe55718b0dcc6bf0334d5af0f6e7

Observation b6b6a7b7-2294-4b70-ba5a-04f1a26dbd1c · inbound

Kernel-based Unsupervised Embedding Alignment for Enhanced Visual Representation in Vision-language Models cites this paper.

Kernel-based Unsupervised Embedding Alignment for Enhanced Visual Representation in Vision-language Models MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:14.425638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:26:14.425638Z digest=sha256:fd410dfe3724fa373c34884d4c001007c1fe99f26dffab8ccf8422b08c2a6ced

Observation dde53622-1195-44dc-bc1e-3d24332b4d75 · inbound

Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills cites this paper.

Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:35:38.573293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:35:38.573293Z digest=sha256:ef79b7868c02904bdf2e33282d3f5f629670144d42df919b0757e9c6c808739a

Observation 3e7cddb7-af68-445e-ab89-eae7961aaa9a · inbound

Neural Inhibition Improves Dynamic Routing and Mixture of Experts cites this paper.

Neural Inhibition Improves Dynamic Routing and Mixture of Experts MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:22:12.586895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:22:12.586895Z digest=sha256:3df24699ae575444606a63ec931a7f79c64037e473e8da57ea3964b3d07976e1

Observation 3edae58d-7dc6-4c2b-870a-7d2a9d41ddee · inbound

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering cites this paper.

CoGR-MoE: Concept-Guided Expert Routing with Consistent Selection and Flexible Reasoning for Visual Question Answering MoME: Mixture of Multimodal Experts for Generalist Multimodal Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:11:53.261984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T07:09:48.239662Z digest=sha256:80c3d4277ab63cc88163dbc9f862ba95262d6d1029cb4a4bd1f7059a04eb3ce6