Pith. sign in

Paper Citation Record · LEDGER

AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2412.15084.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15084 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:33:37.634222Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.516001Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a560f3f3-4675-482a-a41a-9db7deaa3965 · inbound

Process Reinforcement through Implicit Rewards cites this paper.

Process Reinforcement through Implicit Rewards AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:23:30.950535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T20:23:30.763794Z digest=sha256:fd9358703acc01213f35bd18774f8bd76af75b49ef09ce4f80885f20144badd7

Observation d74ba07c-0741-4af1-9acd-5d3812101dba · inbound

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning cites this paper.

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:47:10.237446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T12:47:10.146795Z digest=sha256:41264277c4bed7d5d641bd6104f51870ea59cce17ec823b027729437e59c1398

Observation df7555d8-7f06-4fdf-a615-6036640f0082 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.241571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:61882b76acd98fe4c99f603d6d31620151c2f19e16ec7fcbc5859c29571a3837

Observation a15bbb19-5bf1-4a84-b846-63bc33d71447 · inbound

EasyMath: A 0-shot Math Benchmark for SLMs cites this paper.

EasyMath: A 0-shot Math Benchmark for SLMs AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:37.634222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:37.634222Z digest=sha256:7f35b92cdc36af001303fd4b8518e2707db75ccd58ec1e87f01b128f95e747fe

Observation 943f7d5d-3417-4c90-a906-f22ca245594a · inbound

AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning cites this paper.

AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:35.406551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:06:35.406551Z digest=sha256:011079ddbb94f1fbcfb57439ce638ab8ddf8c2e48cedd002a69018deb7a68d75

Observation 0be86bb8-dca2-4b9f-9e6b-4ee186768982 · inbound

Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem cites this paper.

Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:13:53.714891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:13:53.714891Z digest=sha256:2eb88a9e24a098e999b58eb82ff58158dd4bcb92f5e710392e1eb902c0e5ff87

Observation 30b95365-5af2-40c3-b92b-c5556d29a205 · inbound

Improving Large Language Models with Concept-Aware Fine-Tuning cites this paper.

Improving Large Language Models with Concept-Aware Fine-Tuning AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:28:23.918291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:28:23.918291Z digest=sha256:668165971710c1b1127099d1e65338e2ec464b9ced8faa7d2d92f1edf1371262

Observation d6db3524-6140-4563-bca6-9ab0e975239c · inbound

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy cites this paper.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.455218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.455218Z digest=sha256:13fee49c4ed8a17f43b55b76776fc33b9051f31747aeccc9c75a202800af3d2a

Observation 5b3d8c42-f4b6-4829-a6fe-a0399f6c3e87 · inbound

OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique cites this paper.

OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:11:17.943363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:11:17.943363Z digest=sha256:34c31af244a4e7f19dbd3416ffa22440943120ec6362dea7c61f530b92e2ec72

Observation 0f27a9ea-c53b-4e9c-b200-791d9693c070 · inbound

JT-Math: A Multi-Stage Framework for Advanced Mathematical Reasoning in Large Language Models cites this paper.

JT-Math: A Multi-Stage Framework for Advanced Mathematical Reasoning in Large Language Models AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:11:52.807336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:11:52.807336Z digest=sha256:3ffc6e472732e2605d7cc35f8a05ffc9d882546c66a5b2f2d253374c406dc8d9

Observation 2831f5a5-a1e4-45f5-803c-555773a4587d · inbound

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance cites this paper.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.055041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.055041Z digest=sha256:f3b5ff783489be6ef53bc190e8525fed5d2cfa3fbf0e5ff09e887eff0112db64

Observation 760542d6-a34c-4a94-a573-b9607593e220 · inbound

The Majority is not always right: RL training for solution aggregation cites this paper.

The Majority is not always right: RL training for solution aggregation AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T23:02:32.090666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:02:32.090666Z digest=sha256:a63f6754d35cb96f114146932637193ace57af5edfba82fdec7362c8a8aa8689

Observation 7ad69638-11ba-4189-969a-275c8af045aa · inbound

Fine-Tuning Small Reasoning Models for Quantum Field Theory cites this paper.

Fine-Tuning Small Reasoning Models for Quantum Field Theory AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:24:14.972224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T03:23:18.770963Z digest=sha256:4911365203b40b5f753232ff45ed73d92a2f7133dea17c8a3afd3269f349b543

Observation ade2d6c3-aeec-4041-9d28-85dd780f31e9 · inbound

ZAYA1-VL-8B Technical Report cites this paper.

ZAYA1-VL-8B Technical Report AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 172

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:21:23.382878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T01:15:16.607346Z digest=sha256:6f9aeecad1f2e95ee533f78e6d3b4695a67f5d1479d266a43b1a0e09cdc9739d

Observation bb3f1db2-5ab2-42bc-8764-8d2ea171608a · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 154

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.517246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:09fadb31cff8f16b3f3c0ac4a91c542f274d20cec63012b2384a91e58563471e

Observation f6725e72-865c-4171-8287-ee9018f9808e · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 154

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:219d53d9d80ed425d1c195b5ca9a3335afaf57d77677d1e2a5e6771253166982