Pith. sign in

Paper Citation Record · LEDGER

AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2504.16891.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16891 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 37 of 37 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T12:09:07.287104Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:15:01.137692Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c39e8fbb-1149-40b9-82e9-10a9e0a84215 · inbound

Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model cites this paper.

Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T21:59:02.231426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T21:59:02.173820Z digest=sha256:93faaa32dd1672829b1af14ae7b59f4469f1af3029a3ff8234f624c3f68e436e

Observation da5f633b-de98-4e75-bf85-fc411aeb19d8 · inbound

DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning cites this paper.

DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T10:31:04.827741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T10:31:04.728005Z digest=sha256:3312a5e54d819e4c4a899ced2b5d73e3be32c483963bc6f0b6245f0fd2a573bc

Observation 507bdf62-adfc-47be-af12-419be478a717 · inbound

Beyond Exponential Decay: Rethinking Error Accumulation in Large Language Models cites this paper.

Beyond Exponential Decay: Rethinking Error Accumulation in Large Language Models AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:12:23.951510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T14:07:42.408261Z digest=sha256:0f51494aeaedbbd089112600e5899ade9df9b0f74d9abb432f460144fdbcbd1f

Observation 5508e861-6464-4b36-96a8-358a1acf361a · inbound

Kimi K2: Open Agentic Intelligence cites this paper.

Kimi K2: Open Agentic Intelligence AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T17:49:28.031228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:49:27.926646Z digest=sha256:b8c9a778c3a2f7a2c75dc8cefff28240299cee059276bb0499669bd4e570bf51

Observation dafaa670-7c51-43d9-b42d-e7af38006598 · inbound

Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs cites this paper.

Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T21:16:50.990176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-18T21:16:15.703057Z digest=sha256:a5df3e299bad0f5460b49960b655493f3154ed4dedb1a0749c7aa131b96e331a

Observation 34fd8135-e68c-4d99-9025-f0b1480e7f8a · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:06:13.931204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:d0972392d3676f18170030a41d57e467388811827f651000f1a790eb35490c44

Observation 707f84d1-dbe4-4ba5-a8aa-f6b94c48bcd6 · inbound

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers cites this paper.

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T11:16:00.396506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:16:00.396506Z digest=sha256:f7434c1f14ed9a77f335dc6d91e2a4926144d21df4f59d9dd590632103c1777c

Observation f5c69e64-3f56-4297-8e6e-5f40264c8371 · inbound

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning cites this paper.

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T05:14:19.549625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:14:19.549625Z digest=sha256:2902eadc9df8f679912f0d45b90932e88eb7c8513b3b9656844f7ef7042c1678

Observation 6db84399-daa5-4fcf-ac8a-a3ae41e08b09 · inbound

Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning cites this paper.

Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T21:40:20.875802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T21:38:56.393217Z digest=sha256:a65f8c831ed8078534eaf11fa189fa6e7ee8a92cfb9ae5bb1dcd325cf3f69071

Observation b19025e9-77d4-4d32-a30c-3bc594737b4d · inbound

GLM-5: from Vibe Coding to Agentic Engineering cites this paper.

GLM-5: from Vibe Coding to Agentic Engineering AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:46:41.109901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T05:46:40.836161Z digest=sha256:f387a35a317db8c2482bc87cdd5224413196217877478a7883bf4732b495929e

Observation beada0e5-0295-4b01-a5d9-a832e5328dbc · inbound

On a remark of de Gennes concerning three-dimensional polyelectrolytes cites this paper.

On a remark of de Gennes concerning three-dimensional polyelectrolytes AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T23:55:33.922275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T23:55:33.922275Z digest=sha256:dd1d3fd33eedb5bc6d41bfeec2c23a22dd2d3f719fd0b10b3674897f71429f4e

Observation d4385df7-7038-4a33-abb3-550093c9348c · inbound

OLLM: Options-based Large Language Models cites this paper.

OLLM: Options-based Large Language Models AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:08.956210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T02:04:12.174834Z digest=sha256:9508a0c538f48bcb2fb2450e2c48cb97c25ca5cdccae6af4044ac1a9b1dee6d6

Observation 74564205-abc1-4769-81d5-354f860ac0c0 · inbound

Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation cites this paper.

Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 183

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T03:14:07.779475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T03:13:35.541936Z digest=sha256:6f0ace7c217dcc49e06654e2fbd8ddfda11bb4724298d8394a0f62f6383bb4e5

Observation 458a45dd-d082-4a11-b67d-83bb5a5d5e2a · inbound

Verifier-Backed Hard Problem Generation for Mathematical Reasoning cites this paper.

Verifier-Backed Hard Problem Generation for Mathematical Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:26:08.942219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T11:59:44.850797Z digest=sha256:4b07623a1772abcb335323375345850de523f50c6649e25397a7b29b9fe93a9d

Observation f5e18fbf-c8f4-469d-87cb-fa3d16728099 · inbound

Teaching Language Models to Think in Code cites this paper.

Teaching Language Models to Think in Code AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:15:56.538618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T02:33:02.995360Z digest=sha256:e19c44bb5072c0a240fcd8c24142842f039178fafbdc28767fa6542e1851d5bb

Observation df593bf1-4cd4-4ea9-a732-794bd307ac45 · inbound

Teaching Language Models to Think in Code cites this paper.

Teaching Language Models to Think in Code AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:16:28.264055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T04:26:11.265781Z digest=sha256:0ebd6c7c3748aa7e0c52c362f34e4fe02839d41d49e7ddb3f06c78fa9e803c78

Observation ba8c7ac8-222d-4274-9d8c-2ceded5f2c24 · inbound

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators cites this paper.

Mathematical Reasoning via Intervention-Based Time-Series Causal Discovery Using LLMs as Concept Mastery Simulators AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:55:57.653840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T02:05:48.483640Z digest=sha256:f5d845b96fcf2c632219fc1a1f16ee43e4d796a12d728ab904b2955c6f990141

Observation 003ee557-c391-401c-a978-6b97bd4f3e60 · inbound

GRC: Unifying Reasoning-Driven Generation, Retrieval and Compression cites this paper.

GRC: Unifying Reasoning-Driven Generation, Retrieval and Compression AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:26:26.458524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T03:21:20.024953Z digest=sha256:7ea870306a76fd144820d4936080f3de12125d49ffb0d82881087496bd108dd1

Observation 7354b808-9d85-41ad-98a7-1555f6770313 · inbound

GRC: Unifying Reasoning-Driven Generation, Retrieval and Compression cites this paper.

GRC: Unifying Reasoning-Driven Generation, Retrieval and Compression AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:42:26.060479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T06:41:32.268162Z digest=sha256:439e87fd298395a0af3735b7a713d8eece9755f6d2a0293f16cc0ebb30057351

Observation a02f7a9d-22f3-46ff-8b41-75a19803e652 · inbound

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis cites this paper.

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T21:05:04.665768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T20:58:07.921910Z digest=sha256:9892b9e79289b8e56c39fdb0af0c0b6c205f3f18ea2f40ca2581db16dc200dee

Observation bd7f7d14-0f36-437b-90b3-685c34aa3569 · inbound

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space cites this paper.

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T01:48:28.689307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-15T01:45:39.649473Z digest=sha256:d17a41d27490a258756ae70beeb100aef85de1823ab6f7233bd472e03ad2d9c8

Observation 2d9039d2-2a3e-421d-a5e7-0c78c35ddb5d · inbound

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space cites this paper.

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T16:32:39.576794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T16:27:59.293776Z digest=sha256:1a94ebbc2e8c0f22dab47e3961962d786f648e5e25d4f7b4646a8e136a659d6a

Observation 65cf7f42-4011-4d67-93a6-863c80099034 · inbound

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space cites this paper.

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T14:35:46.486098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T21:16:13.804718Z digest=sha256:94fdc72e348fc9b3d88bc581e7ce5285d982bf313d8d1da292803dd5f75a014f

Observation c3f5ad27-c804-4b4a-acff-a196a9bfc9e5 · inbound

Unified Data Selection for LLM Reasoning cites this paper.

Unified Data Selection for LLM Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:34:40.102955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-22T05:33:20.930156Z digest=sha256:a5871e3019f7b801c26958bde112b33524b79500486d2b8337e3521615a82a53

Observation 988e712b-ec23-442f-beb0-7f45f1678da6 · inbound

Not only where, But when: Temporal Scheduling for RLVR cites this paper.

Not only where, But when: Temporal Scheduling for RLVR AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.841016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T22:34:44.791173Z digest=sha256:446314fdc63c350470118f5cb6c1ddfde29a5d4d2c3217c7d4e8c71792182cd0

Observation d79671a2-5805-4944-bf0c-ac46caac62ce · inbound

ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation cites this paper.

ReverseMath: Answer Inversion for Scalable and Verifiable Mathematical Problem Generation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:03:47.908828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T18:01:05.196759Z digest=sha256:9e786fca864a96105d3d70976513fb66e2e8e2616b2519937c029a50680f1733

Observation 1cc22062-26ed-4fea-a483-97beffd55d5e · inbound

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL cites this paper.

Extrapolative Weight Averaging Reveals Correctness-Efficiency Frontiers in Code RL AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:43:30.584338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T14:41:13.191919Z digest=sha256:ff837ab35b14cb993ceebd672e17cb95f6da90f684356b42ef6be6f4f83a75aa

Observation 0eefff4d-0632-4742-9d31-82d378465075 · inbound

Value-and-Structure Alignment for Routing-Consistent Quantization of Mixture-of-Experts Models cites this paper.

Value-and-Structure Alignment for Routing-Consistent Quantization of Mixture-of-Experts Models AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-28T01:31:28.819320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T01:29:20.772523Z digest=sha256:65e32fd8acd41d4fed66cb35afc4244689922e6247f603987b50bf1cfde13165

Observation fb2d72d1-bb01-437a-be1e-8406d51f0b45 · inbound

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning cites this paper.

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-03T09:17:48.658520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T10:26:51.429436Z digest=sha256:ca12f9543de3a55efe20413c9fddc5270bae8e63bcc253abde7bbe5f89f12935

Observation 04242b9a-6f4f-4bcf-9f87-1a28cf233536 · inbound

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning cites this paper.

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T11:51:35.543118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:51:35.543118Z digest=sha256:085ef07032d2bb7502b2aea67131eb02cfffcbcb14f9d43f05f12db1cd121022

Observation 05db2b43-7daf-473e-8ae1-2959cebd784e · inbound

Fine-Tuning Large Language Models for Quantum Reasoning cites this paper.

Fine-Tuning Large Language Models for Quantum Reasoning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:09:41.917454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T12:02:14.680564Z digest=sha256:b06ac72d2846ea93ba907dcdf5e5188f5da598d46fa7c151b36e2060a3141506

Observation 8e1c242d-f7c6-4b86-8ac5-633c511f1db9 · inbound

EntroRouter: Learning Efficient Model Routing via Entropy Regulation cites this paper.

EntroRouter: Learning Efficient Model Routing via Entropy Regulation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:44:21.626373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-30T07:42:35.962864Z digest=sha256:7ae97a6e8565b3ff734a8a25c3f6a28ec4db5b1cde056d7c376f9ab880ff0a64

Observation 34cfb234-605f-4a03-9167-77d38478cf09 · inbound

Don't Let Gains FADE: Breaking Down Policy Gradient Weights in RL cites this paper.

Don't Let Gains FADE: Breaking Down Policy Gradient Weights in RL AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T21:08:57.782401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-03T20:59:57.539909Z digest=sha256:2a77ab0e0c8dc1b47ee123c36452d12f4989c898586b0129cb21c4dc806c5a97

Observation 40f62368-9b92-464a-bbda-bdcc15dd56e6 · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T12:33:45.121834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-07T12:31:42.224094Z digest=sha256:944efd2f7a8e645338fb447f53bbd96188089ea1b8adb251ee963990ae79a78b

Observation 75e7e64b-cea9-457a-b420-8432ff36543f · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-11T07:01:56.628017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:01:56.628017Z digest=sha256:8a6f7e64fc18ba56c801aa645153b9352c0d2a7deb14e7c1e91ac5a38cfff08d

Observation 8c69147c-83f9-4f0f-9f52-e72e4136b657 · inbound

On-Policy Delta Distillation cites this paper.

On-Policy Delta Distillation AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T00:00:06.740351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:00:06.740351Z digest=sha256:6a2c2e00b09138eca0d11f3dd52f27032b0c527d5de34caeac6266612031bc1d

Observation 7d56f8fe-1f77-469b-b8b3-0e315a405ce6 · inbound

Learning Latent Reasoning Traces for Scalar Reward Models End-to-End cites this paper.

Learning Latent Reasoning Traces for Scalar Reward Models End-to-End AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T12:09:07.287104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:09:07.287104Z digest=sha256:4bfef63bf0d463305ab6c812982197497c5b622e1447101716d28f1a376ef753