Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2505.04623.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:13.645079Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T06:36:10.581673Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation bb8cd184-7364-4401-a920-23f1239dae01 · inbound
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8809531c-3f7a-4032-8af3-4ab46ef679fb · inbound
FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2caf0f52-2c2f-4cb8-bcf5-c365bbbc7f9b · inbound
HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f29198a-1efa-4caf-8739-a82b313520b7 · inbound
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 264
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0a7d057-b8c4-4e62-9dde-6caf0ffe500f · inbound
XModBench: Benchmarking Cross-Modal Capabilities and Consistency in Omni-Language Models EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6cbf7b30-a75b-4436-97f2-cf6cb8b1392d · inbound
Development of a 3D-CNN-based Prediction Model for Migration Barriers in Plasma-Wall Interactions EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9dfaff-cf0c-4434-995f-49833a1bc469 · inbound
Cross-Modal Coreference Alignment: Enabling Reliable Information Transfer in Omni-LLMs EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb2b7095-f2cf-4861-8605-52d8989c1bad · inbound
Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87eebf59-fac5-45cc-b15c-4399c9236ab6 · inbound
Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1afb3801-19f9-4cf1-a99e-336a9f694b0a · inbound
Chain of Modality: From Static Fusion to Dynamic Orchestration in Omni-MLLMs EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43e2786f-801b-4e1b-aab3-a201e66cfbbc · inbound
AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dcd6a649-c1a0-4046-af9f-6f01963efd5f · inbound
LatentOmni: Rethinking Omni-Modal Understanding via Unified Audio-Visual Latent Reasoning EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.