Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2504.10462.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:53:05.867758Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T13:33:28.015082Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 97348b64-4114-4889-9a2a-b871dc941f75 · inbound
Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb0a40c4-3a26-4ea0-bf95-5f2fffdf4283 · inbound
VGR: Visual Grounded Reasoning The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e3d7bf0-8038-4b4e-9817-32ed004715e8 · inbound
SAILViT: Towards Robust and Generalizable Visual Backbones for MLLMs via Gradual Feature Refinement The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42056ed7-d647-444e-86bf-c2139b686188 · inbound
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9299fd8b-3ac9-4be1-b289-3704a39184a0 · inbound
Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3dcd580-27df-4fee-9729-3b3de8ca34c9 · inbound
From Pixels to Words -- Towards Native One-Vision Models at Scale The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 458afbc7-108a-4f0e-9017-244ed6609077 · inbound
MuRA: Multi-Rank Adaptation for Efficient and Effective Test-Time Vision-Language Generalization The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single Transformer
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.