Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2104.12756.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:17:27.900588Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T00:49:18.047840Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c16c13f3-b74c-486a-8ff6-11fc527d18b9 · inbound
OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning InfographicVQA
Reference 119
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cbbfbcab-553b-455a-a67e-1aded7501099 · inbound
Survey on Question Answering over Visually Rich Documents: Methods, Challenges, and Trends InfographicVQA
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81917d7c-899a-4039-8374-d040546404e5 · inbound
FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding InfographicVQA
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 27d21c87-cfb0-4cf0-9ce8-3a1f3b7b226f · inbound
Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models InfographicVQA
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ec357329-5c77-4802-9e4d-91fd25aa5219 · inbound
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models InfographicVQA
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc315a0f-de32-4254-b33f-c504ef33e642 · inbound
IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation InfographicVQA
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8f2ebf-1202-40bd-8ece-d5a17c6b85cd · inbound
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 060bf5c0-ab02-4e9d-8812-72eed8a0eecd · inbound
Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning InfographicVQA
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6977d310-e376-4565-881f-16f8d8fec6cc · inbound
Chart-RL: Policy Optimization Reinforcement Learning for Enhanced Visual Reasoning in Chart Question Answering with Vision Language Models InfographicVQA
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3a7fc88e-2bdf-4c7c-ab6e-45ca67073d3d · inbound
Vision-Language Foundation Models for Comprehensive Automated Pavement Condition Assessment InfographicVQA
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a517b7d4-85fb-4cae-84c7-cb1cc995c878 · inbound
Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models InfographicVQA
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b10619f3-6942-44e7-afb2-1f900912b75c · inbound
Enginuity: A Dataset and Benchmark for Vision-Language Understanding of Engineering Diagrams InfographicVQA
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2a552759-cfe3-49c6-a22b-78088ce85eb0 · inbound
MODE: Modality-Decomposed Expert-Level Mixed-Precision Quantization for MoE Multimodal LLMs InfographicVQA
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5a9c32be-0c02-4617-93f7-9729fd6f1cfb · inbound
PerceptionDLM: Parallel Region Perception with Multimodal Diffusion Language Models InfographicVQA
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c35469fc-3a31-4e53-a0a4-01e99de5bbe6 · inbound
Combating Textual Noise and Redundancy: Entropy-Aware Dense Visual Token Pruning InfographicVQA
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 42c284ec-8a14-49c4-8d1a-f6c3308df46d · inbound
Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model InfographicVQA
Reference 127
Source-reported events for the cited work
Unavailable: canonical work link unavailable.