Pith. sign in

Paper Citation Record · LEDGER

Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2505.20256.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20256 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:07:21.383338Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:27:15.771664Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 748bcc57-ef43-4c6c-874a-623cafbcc7f9 · inbound

Group Relative Policy Optimization for Speech Recognition cites this paper.

Group Relative Policy Optimization for Speech Recognition Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:21.383338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:21.383338Z digest=sha256:d64729065cc8b7a7e8a2b1217de37176d83f45c68fb24b804a7238a68c4b0f4c

Observation 11fa81bf-7cdf-4185-802d-a848fa6ccfb4 · inbound

XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models cites this paper.

XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:45:56.128382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T05:45:07.700571Z digest=sha256:791f82af8ebde1dce6319bab6c24a45b0db9cba343dbe93b29a24ec8a76d13ce

Observation 9a349702-83c9-405d-ae4c-4abb8a12968e · inbound

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling cites this paper.

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-22T12:31:32.120343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T12:26:35.347190Z digest=sha256:99db9b732fa8b176c9af7836d81d568e31abd19b054abcb47c2eeed3c4b0232e

Observation 60a8a706-ac62-4312-9138-576e90d3bfca · inbound

Omni-R1: Towards the Unified Generative Paradigm for Multimodal Reasoning cites this paper.

Omni-R1: Towards the Unified Generative Paradigm for Multimodal Reasoning Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:37:59.982181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T14:37:05.402850Z digest=sha256:d915d1f5da2b929fcae7a296878a1da201dbae2d14a5791addaca60a93f70069

Observation a35bda6a-7128-40f8-b307-a518e00bcfc2 · inbound

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering cites this paper.

OmniJigsaw: Enhancing Omni-Modal Reasoning via Modality-Orchestrated Reordering Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:11:01.973351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:45:51.528645Z digest=sha256:03a8c691c132a7e162d098c20d272ed9f497edd89a1e697874e3ee7a00192d62

Observation 5cd62825-00f7-443b-b17b-2f24b32fc7a8 · inbound

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding cites this paper.

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:11:03.827672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:07:45.595260Z digest=sha256:98ce8042c0a1827b64a880d10fc1d84085f0d48f9f9d0622a3c323370308e94c

Observation 3eb7c43c-ccbd-4e9a-a055-c0581c6ba1b4 · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.298818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T14:10:03.707886Z digest=sha256:f22e87cb5bbacbb0e48ffb5732b554b8438a4e3820ced3aa530090df74d7273a

Observation 6eb1a9ba-ff0e-40fe-a774-7737a268bd6a · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T21:18:46.566338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:18:46.566338Z digest=sha256:c55a0b7b6226cc317fb70c8135d623f0e132d42f7c40984c325587cba42ee2ba

Observation 33213efd-ece5-4758-826a-c2693ae4e7c5 · inbound

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs cites this paper.

Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:10:22.082628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T12:05:54.551728Z digest=sha256:d01df7e0eb40a28f73a07a6d54b096409ace3061a11bcb05b7baa07b26c16dd3

Observation 6baa4af0-3229-4b99-b6c1-dfea506983bf · inbound

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers cites this paper.

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:18:32.211166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T08:02:53.574120Z digest=sha256:ac2502ea06eac4d3f143c6dec3724434fab06c09d1a78422d51ec19138c02ad8

Observation cf8a1670-c68e-4524-ad51-1587028d1389 · inbound

PRIMED: Adaptive Modality Suppression for Referring Audio-Visual Segmentation via Biased Competition cites this paper.

PRIMED: Adaptive Modality Suppression for Referring Audio-Visual Segmentation via Biased Competition Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:30:53.325802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T02:30:44.486435Z digest=sha256:7ae3485dce648d1798a5a009a7735abece0b7b08f845f596eef570c04eb0358c

Observation e46a2ae7-4b41-41ec-9c82-fcdb79ddfe19 · inbound

RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation cites this paper.

RCoT-Seg: Reinforced Chain-of-Thought for Video Reasoning and Segmentation Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:59.904231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T01:16:25.031349Z digest=sha256:592fc7a83c20f0821dbfe3d70150044cdcb95dc54b4a691bb90a7b5757aeee03

Observation 59631231-82f1-48ef-91b3-68a6c3106ccb · inbound

Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching cites this paper.

Eliciting Complex Spatial Reasoning in MLLMs through Wide-Baseline Matching Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:36:27.106597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T10:48:24.523702Z digest=sha256:8d6ebe75ec8e031a982cdccacd41bace9647358566949876453c922991a4c91f

Observation 6c937d94-f391-4895-86bf-b614ce2e5c02 · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs Omni-R1: Reinforcement Learning for Omnimodal Reasoning via Two-System Collaboration

Reference 202

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.773107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:d53d329b530b8f4d92f993e616900d5f9083b545af0c33ca7c8ec5c42783552c