Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2503.10497.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:44:35.335823Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 0cfa8bad-7bbd-4edb-9bd0-d17271414661 · inbound
CapBencher: Give Your LLM Benchmark a Built-in Alarm for Test-Set Overfitting MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4a59685-f744-4586-8b85-b098e6b5a1e0 · inbound
MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b62f5b2a-4507-44ed-9183-e799adadf83b · inbound
Do LLMs exhibit the same commonsense capabilities across languages? MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bf650b9-a6a6-4ba8-b0e1-e216529eed56 · inbound
Why Do Multilingual Reasoning Gaps Emerge in Reasoning Language Models? MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bec86f7f-e4c7-42ae-b56a-a0c52930a7e2 · inbound
Is Biomedical Specialization Still Worth It? Insights from Domain-Adaptive Language Modelling with a New French Health Corpus MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 6239d49b-95bd-4115-bc44-72ccd4a0f39e · inbound
COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 91014f22-4b00-4901-a80a-72d5e40d0df2 · inbound
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c67d6216-ad29-4182-bc2c-ddc9650365db · inbound
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 80e1b972-e75d-400f-a470-081157ee5f03 · inbound
Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 99c475e1-693c-43a0-9268-58784579774e · inbound
Creating Multilingual Mental Health Dialogue Datasets: Limits of Persona-Based Localization via Nationality and Language MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a5ab1097-0f80-4c81-9db2-e2c56a476596 · inbound
Disentangling Language Modeling and Boundaries MMLU-ProX: A Multilingual Benchmark for Advanced Large Language Model Evaluation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.