Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T09:09:14.884516Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 79 inbound Pith citation observations for arXiv:2410.07985.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-15T09:09:14.884516Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T00:22:29.920179Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
77 of 77 outbound references displayed
External citation measurements
3
pith, observed 2026-08-05T02:28:24.338817Z
Observation 111cc7e6-2451-4361-aa0d-ca88ac5a2b0e · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2021 , eprint=
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e9aeaf62-f410-429f-9999-35287294270d · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2021 , eprint=
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3af6d83b-34fc-4a7a-85f1-b2e1da473f04 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , eprint=
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 505469b3-8300-4c43-842f-44e6e7ce5d39 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2022 , eprint=
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 589dd251-4d28-460a-a33e-6d168d24bb01 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , eprint=
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c6d9f78e-e229-40f4-896a-0c92ffbe1de5 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7b13af36-1990-4106-9565-642bbbdbed9d · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9e155c53-2963-40f3-94ab-6a628b2580f0 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3c47b969-6e11-404f-b692-08b0e7f5e12c · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 12844f27-7b36-4e0f-a029-e27d43e9de05 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2022 , eprint=
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 642ec3fe-198c-440f-b011-7be415e9772f · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Nature , volume=
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 50dc1e93-050b-4592-bd36-586791f598f7 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Hugging Face repository , howpublished =
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d485c82f-a1ce-4ae3-b53c-af0ba529e8ff · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , journal=
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e3d42753-0873-4f37-bad3-b09685b7b503 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , eprint=
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f52c174f-8e6a-4658-9266-560fa96a752f · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , eprint=
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 754cb1f7-23a0-479e-93f7-77da931dd116 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 366d2d5f-a2f8-4c43-a123-f5d006708522 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6ea17305-02fe-442e-b22d-4cd4e769fd1a · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Journal of the American statistical Association , volume=
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f5c263ef-df71-4895-b674-4f1e1af72780 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , url=
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation adddf800-f79b-437f-94f4-5494c36538ea · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a3a10510-fd94-4695-a29b-e60e04daffcc · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 706739d1-378b-48fe-9511-b442ee8857c0 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bf6928b8-9e69-447e-9d7a-06c97cde6188 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 57804fda-d753-418c-8e5b-7738767212dc · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0a304a85-e77f-4e5f-b010-f4b0c4fe4440 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5af3e1bf-e4e6-44db-a025-42db2db54a46 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 43125d1c-8e83-489a-95e5-c2b4f1dd63b5 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6d6e929d-b0dd-47b2-98d3-d2aad9758169 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , eprint=
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 249eeee5-f771-4c02-9941-a345786ecd45 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 93c1207e-74cc-44d3-a4ae-0b2981a7c2a5 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f3cf1fc6-0ed7-4a76-877b-80afac2210a7 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2023 , eprint=
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8d83c05e-14f3-4525-ae54-d39510b7ac8d · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models An Iterative Optimizing Framework for Radiology Report Summarization With ChatGPT , volume=
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9fa867a5-85c0-4392-8a3f-bd22735b11e9 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 09c0cc3f-75bd-41ab-aa35-54d484f05674 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7a97d095-6e60-46b8-8736-271fc0467445 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ad6e1c53-3095-41f6-965e-f75d0ed9d5e1 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 756c2935-d152-4b32-9643-c94dc7c26072 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 02aa4479-120f-4bd4-9a46-c054b88a38a9 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2017 , publisher=
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4e1d2122-6d4f-44de-b677-8d883edf776c · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2c09868b-e5cb-401f-8cff-d0ebf2cb2936 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , eprint=
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0af20043-a3a2-47ad-833c-293a8b067571 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , howpublished =
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a776d289-69a4-4f03-ba6a-fe9c4860b581 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models 2024 , howpublished =
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c9de084d-101d-42eb-bbe8-1a268c138f4f · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models The Llama 3 Herd of Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c6123048-fa25-469c-8661-199bd963979f · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Mathstral
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6c182db5-ea1b-4417-be5b-c356e81147b9 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Claude 3.5
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation dcdd4251-351b-4034-a861-a2f72e6272b3 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Have LLMs Advanced Enough? A Challenging Problem Solving Benchmark For Large Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 718c13df-d833-4de6-92f7-37674a21ecbe · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models ProofNet: Autoformalizing and Formally Proving Undergraduate-Level Mathematics
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 38fd9800-16e8-4cc8-bc16-d25f807c602d · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Llemma: An Open Language Model For Mathematics
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 345a8872-941e-40c6-a086-813961ca63e8 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Distribution of residual autocorrelations in autoregressive-integrated moving average time series models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 200f367c-3777-429e-bee8-0ac5ee91faeb · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 41b902db-e93b-49dc-95fa-b5941b896d78 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Training Verifiers to Solve Math Word Problems
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 55910282-9409-4c43-9999-233ff335fcb5 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9ca48e62-59a7-42c7-a127-2ce6a7806efe · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Problem-Solving Strategies
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ed9b2a01-ae68-458c-8954-79d7f91a31ce · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 174bd1a2-1c24-4758-98fe-7cd4aa632bd7 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models LLM Critics Help Catch Bugs in Mathematics: Towards a Better Mathematical Verifier with Natural Language Feedback
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 36ffdc63-201c-4d1f-88d0-2d8df8affe94 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 089a28c4-3dbf-4bf2-ba7f-979b43201aa9 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e1fe4b74-821d-4190-b2de-d826281acc9a · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 02d4b553-6a98-4c4a-9895-575d3465edef · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Qwen2.5-Coder Technical Report
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b936ff09-3887-48b4-9ff5-c3e429fdd334 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Solving Quantitative Reasoning Problems with Language Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f606ce41-2851-4a0d-a558-cd6fd85cf1f8 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Numinamath
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation dffef66e-f120-4612-9347-99c579b7279a · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e9627305-6f7f-45f6-9fe3-6b5438bff278 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Gpt-4 technical report
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 67f2e29f-5fbb-41e7-a92b-897233ffd297 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Learning to reason with llms
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cc2dce73-a10a-48bd-ade3-79fe29458d89 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Code Llama: Open Foundation Models for Code
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 372f2faa-7aed-4b5f-8ab0-6b23354cbcab · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8a75b2ab-cca7-4914-ac35-87d091de9a1e · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Solving olympiad geometry without human demonstrations
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 817fd360-ae84-4c74-a0cf-b7d6cf23ac86 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5500fd44-6682-4a2d-8452-5aca098e8a37 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Preserving in-context learning ability in large language model fine-tuning
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 35845f36-317b-421a-94c2-f2807d182c79 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Benchmarking Benchmark Leakage in Large Language Models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 136e4829-a1bb-4adb-b945-fd2378d31e71 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Qwen2 Technical Report
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4c88fc30-871c-45d3-a12d-950f9e21c535 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f8baea53-1534-4f46-8b4d-0b2b7d06328e · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models Can Large Language Models Always Solve Easy Problems if They Can Solve Harder Ones?
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ee60e291-cb29-4f52-9319-9e737ce448b5 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3e027ceb-a0f1-4b9e-8ad4-8990c7b47dd4 · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models The art and craft of problem solving
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 25abb680-ff6a-40a1-8d0d-78b693ab81ef · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 218aff1b-80eb-48ba-8f6f-138e2144ab7d · outbound
Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models MiniF2F: a cross-system benchmark for formal Olympiad-level mathematics
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 15afcf64-1beb-4bcd-9a12-3cb633a8a7d3 · inbound
Confidence v.s. Critique: A Decomposition of Self-Correction Capability for LLMs Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30800675-fe33-4030-8767-e9e2aca13a71 · inbound
CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 1978
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bb0e081-ff39-4c5d-9db7-e4aa955c24ef · inbound
End-to-End Bangla AI for Solving Math Olympiad Problem Benchmark: Leveraging Large Language Model Using Integrated Approach Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6f7cb30-a856-48c3-b2c4-36aabad2993d · inbound
T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e914fe3-e237-4b2a-9699-84239dbf0294 · inbound
Humanity's Last Exam Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 12a7c8d0-1ed8-4fc8-82b8-03ae4d6ffb86 · inbound
UGPhysics: A Comprehensive Benchmark for Undergraduate Physics Reasoning with Large Language Models Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 906baf4e-c8d2-41ed-9100-66a0d22d47f7 · inbound
L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a6b304e-0c60-4719-85e4-d978156008d5 · inbound
LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b94038d3-60fe-489f-aee2-ede6b1573753 · inbound
MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 91278f42-2e3b-46c9-a8e9-461d97f33d26 · inbound
Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 85cff00c-7d90-46d5-a7ff-42656da1cbae · inbound
DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2e341750-2138-46c8-9ec1-da57a5347267 · inbound
Reinforcement Learning for Reasoning in Large Language Models with One Training Example Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3a56fac1-4f09-493f-ad1f-d88b1ad02e51 · inbound
The Hallucination Tax of Reinforcement Finetuning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67a7076-9d68-40a0-920b-580c6667ec10 · inbound
SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae6bea58-246b-4a2f-afdb-9148b5a5af40 · inbound
Incentivizing Dual Process Thinking for Efficient Large Language Model Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 887e1314-3873-42ed-9526-052b06796094 · inbound
AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31ca3bd3-192a-4723-a0f4-dd4082dc77b7 · inbound
RBench-V: A Primary Assessment for Visual Reasoning Models with Multi-modal Outputs Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ee06ac6-e9c8-4669-a56a-d1e83aa8f32b · inbound
Thinking Fast and Right: Balancing Accuracy and Reasoning Length with Adaptive Rewards Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c3b15ba-1857-43f8-b4b8-5ab6c984caad · inbound
Formally Solving Answer-Construction Problems in Lean Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb9f2095-3250-48b0-8e20-4e991cea372b · inbound
CoTGuard: Using Chain-of-Thought Triggering for Copyright Protection in Multi-Agent LLM Systems Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8570983-096a-48d2-96c9-cb11b5ddd572 · inbound
Walk Before You Run! Concise LLM Reasoning via Reinforcement Learning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69415067-5e5a-4773-98af-3df7ece59e28 · inbound
Decomposing Elements of Problem Solving: What "Math" Does RL Teach? Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07659698-598a-4887-90eb-6701309475c2 · inbound
MathArena: Evaluating LLMs on Uncontaminated Math Competitions Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 35e7b365-250a-4634-b96e-a7e534e7b8db · inbound
Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e711829c-0ac3-47de-a042-751ba8441949 · inbound
Reinforcement Pre-Training Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7d0511f-bc7d-4b3f-9910-c310f4ad58b5 · inbound
TreeRL: LLM Reinforcement Learning with On-Policy Tree Search Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c5f84c5-589f-4ffe-9226-2f2ca8fa44f7 · inbound
SciDA: Scientific Dynamic Assessor of LLMs Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e2253e-bac5-43a5-9fd5-69693d221042 · inbound
CriticLean: Critic-Guided Reinforcement Learning for Mathematical Formalization Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa384924-2e54-4492-a8df-de4af4faa6ab · inbound
One Token to Fool LLM-as-a-Judge Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32029885-2627-44ca-a02e-fa8a0510044d · inbound
Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06093aef-6291-4c7e-aa5b-3a6027b33ad1 · inbound
REST: Stress Testing Large Reasoning Models by Asking Multiple Problems at Once Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c4b594e-f0a2-43cc-8651-0f6f4b6a0a1d · inbound
Supervised Fine Tuning on Curated Data is Reinforcement Learning (and can be improved) Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 400cd132-617b-436b-a6b7-43a1cc4c6d71 · inbound
Proof2Hybrid: Automatic Mathematical Benchmark Synthesis for Proof-Centric Problems Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46efc729-4bcc-469c-9444-c125fa14dbf2 · inbound
Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3ec06ee-2212-451f-a16d-175cb0800d5e · inbound
Throttling Web Agents Using Reasoning Gates Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98543ed1-8fd0-45c8-9acf-285ac669280c · inbound
Another Turn, Better Output? A Turn-Wise Analysis of Iterative LLM Prompting Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4aca38e9-be89-4ee7-a789-2e9140c49eef · inbound
RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c49559d-194a-47fc-bfaf-fecdf23671ac · inbound
Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfd12c5a-cb70-4bee-8fb7-e4975cf227cf · inbound
EngiBench: A Benchmark for Evaluating Large Language Models on Engineering Problem Solving Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c4988a27-d134-48c3-8ddb-236d8af9ec9e · inbound
StatEval: A Comprehensive Benchmark for Large Language Models in Statistics Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06dfca44-1c00-4d8c-8ece-9b7bb9ab5c00 · inbound
MENTOR: Reinforcement Learning via Flexible Teacher-Optimized Rewards for Tool-Use Distillation Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17d1b9ca-423f-438e-a586-6552ca628b97 · inbound
LLaDA2.0: Scaling Up Diffusion Language Models to 100B Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7b1cd48c-cbdb-468f-9303-5e6e6d6268d3 · inbound
The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e0d579c9-2615-42c2-8439-a081d552c22b · inbound
ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dd39d68-61f9-42c5-a67e-6b7a1480586d · inbound
LemmaBench: A Live, Research-Level Benchmark to Evaluate LLM Capabilities in Mathematics Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab5b8ce9-1ab1-4587-93b8-4af6cc0e48a2 · inbound
OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db0b71ae-c859-4823-9861-72ff92aa634a · inbound
Open, Reliable, and Collective: A Community-Driven Framework for Tool-Using AI Agents Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a524c0e0-bb4c-4317-bf1a-13eaff005a87 · inbound
Riemann-Bench: A Benchmark for Moonshot Mathematics Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 282951e6-a92a-4ea0-b2db-4fbf7f691e22 · inbound
HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 93bd2711-d30d-48f5-afb6-351a5c074111 · inbound
MathNet: a Global Multimodal Benchmark for Mathematical Reasoning and Retrieval Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 838ade7f-4b5b-4def-8cd6-70bb4f0a10e2 · inbound
OLLM: Options-based Large Language Models Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f1c004be-7c19-4585-9da1-f9a2e6d475f6 · inbound
Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4cd45cc3-8a0d-497b-a96f-578296a69bb5 · inbound
OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ca08261e-a418-411e-ba16-d09ca7c1cfa7 · inbound
Rethinking Math Reasoning Evaluation: A Robust LLM-as-a-Judge Framework Beyond Symbolic Rigidity Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 282031cd-6c11-4d82-b87c-0e9cc8b09057 · inbound
Uniform-Correct Policy Optimization: Breaking RLVR's Indifference to Diversity Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation df0b4dbe-0944-4a03-9b9b-33ac3c1ab829 · inbound
Verifiable Counterfactual Supervision for Process Reward Models Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9bc3e4fd-85be-4e4a-a740-1c965bd41702 · inbound
You Snooze, You Lose: Automatic Safety Alignment Restoration through Neural Weight Translation Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cb507dec-0fa6-4004-94fb-048b23aa4fc6 · inbound
Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 53ef6107-7bd0-4e64-a5f9-2b8924308ecf · inbound
Iterative Critique-and-Routing Controller for Multi-Agent Systems with Heterogeneous LLMs Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fdf3d0ef-00a8-4fef-aa75-ada9cc889f2b · inbound
Beyond Accuracy: Evaluating Strategy Diversity in LLM Mathematical Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation af13ff09-c550-4d0e-98ec-439e0bb4be80 · inbound
TIDE-Bench: Task-Aware and Diagnostic Evaluation of Tool-Integrated Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation cc48f11c-e718-463d-919b-7b742f5c0ecd · inbound
Nice Fold or Hero Call: Learning Budget-Efficient Thinking for Adaptive Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e12b708f-6244-40cb-8a63-c6f760dcef65 · inbound
Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalism Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 362cf041-0fe3-45e8-93cd-7629f955998b · inbound
RMA: an Agentic System for Research-Level Mathematical Problems Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4b75cfb3-a1ca-458a-bf81-03c02e72d9f7 · inbound
SkillOpt: Executive Strategy for Self-Evolving Agent Skills Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f66cc439-9b5f-4175-9773-67ea13d92506 · inbound
SkillOpt: Executive Strategy for Self-Evolving Agent Skills Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 90451569-2903-46be-946b-89165f4f4bdc · inbound
RLVR Datasets and Where to Find Them: Tracing Data Lineage for Better Training Data Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 17200321-6e33-417a-8953-6feca06a231d · inbound
Bridging the Detection-to-Abstention Gap in Reasoning Models under Insufficient Information Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 377fb678-6ebb-41f5-a322-7c1f5ba8d447 · inbound
Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 219
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 38de89c2-d281-47f0-ae08-e4f395a9ff85 · inbound
Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 221
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79c3b2ca-0396-4c14-ae59-cdf5706853f7 · inbound
KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2aecc018-1a2a-4a5e-8c03-b6b612e973a6 · inbound
Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b4d53ea3-6e0a-436d-82fe-6f0fc22e85eb · inbound
Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 604b69a1-63f9-41af-9cbd-e914229b0396 · inbound
Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 54ae2a75-170f-4185-8607-c99e59f0dca6 · inbound
OS-Pruner: Pruning Chains-of-Thought of Reasoning Models via Optimal Stopping Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96536195-28b5-464e-870e-60c9d5909ec8 · inbound
AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f3ef200-f229-4067-a7f3-498e6f885606 · inbound
Precise but Uncoupled: Reviewer Precision Does Not Guarantee Critique Uptake in Multi-Agent Math Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95d6c613-43c4-4846-8694-4ff07d2545a9 · inbound
Off-Context GRPO: Learning to Reason on Hard Problems using Privileged Information Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4263efc-e780-4a88-91e7-e55119155782 · inbound
When RLVR Shrinks the Reasoning Boundary: Diagnosing Pass@k Inversion Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.