Pith. sign in

Paper Citation Record · LEDGER

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 27 inbound Pith citation observations for arXiv:2506.13284.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13284 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:41:36.479613Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:52:33.396332Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 348a1514-6e0a-4921-924d-a45253850050 · outbound

This paper cites OpenCodeReasoning: Advancing Data Distillation for Competitive Coding.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy OpenCodeReasoning: Advancing Data Distillation for Competitive Coding

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.663663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.663663Z digest=sha256:241b785f26ba3c6cba7ab39278f62ff88edcf26b8e7fa520c3d8eae084a20baa

Observation 251cd77c-ad37-4a0f-9188-74970e050770 · outbound

This paper cites Matharena: Evaluating llms on uncontaminated math competitions, february 2025.URL https://matharena.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Matharena: Evaluating llms on uncontaminated math competitions, february 2025.URL https://matharena

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:40.297557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:33.722383Z digest=sha256:7b6723414366b630761e373782b6b0395634e33bb43d27caad849e7ecde44f5f

Observation 9854e4b3-d6fa-48ee-b962-b75be4457496 · outbound

This paper cites Llama-Nemotron: Efficient Reasoning Models.arXiv preprint arXiv:2505.00949, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Llama-Nemotron: Efficient Reasoning Models.arXiv preprint arXiv:2505.00949, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.785649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.785649Z digest=sha256:ed8c76cddd07d72522c6b04b95cc35106fc7a90ff4f769c90cd4df1b1d3be2e7

Observation bdb8ca73-c91c-4ec5-a086-5b469694843f · outbound

This paper cites Evaluating Large Language Models Trained on Code.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Evaluating Large Language Models Trained on Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.875463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.875463Z digest=sha256:fe1fbceedba74d7f75eabeb42dbe5a91cc3cd84c9afb03281dd7795ebe74dacd

Observation c2edf7f1-ed2c-4a8f-80a6-07974803d571 · outbound

This paper cites AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:33.940080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:33.940080Z digest=sha256:abfae022dd0f40b3c49d87f81940831255ddf7a694a896134a6829dd1dac7b36

Observation ddb7e23d-d735-4ec3-a22f-3d838b10963b · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.004261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.004261Z digest=sha256:813e85803dca100a1cfe210ab5573054fb48a23b06f0dfd32f6b7f3efdc8195c

Observation d4944494-ca7c-4a38-8527-d901cc9d6985 · outbound

This paper cites NVLM: Open Frontier-Class Multimodal LLMs.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy NVLM: Open Frontier-Class Multimodal LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.085376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.085376Z digest=sha256:3bbda9c8236c9eb341f6381afe3a74a637e5bdd405a90a97d0f2e714605f66a3

Observation d6eb32c1-0faf-4999-86af-f60a5b54c463 · outbound

This paper cites Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Audio Flamingo 2: An Audio-Language Model with Long-Audio Understanding and Expert Reasoning Abilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.161298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.161298Z digest=sha256:d4de79e2e1024bdefa6b22e799f6a178bca8b8657c94495d40f4a561e66c51d3

Observation 19f5eb4a-102d-4841-ac28-983ecd9d139c · outbound

This paper cites The Llama 3 Herd of Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.233620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.233620Z digest=sha256:697302c29a426dca970c49a3fdc3644be59f4efd7df2f0e8055bf38e7e2c04a8

Observation 11c2f08d-8cfd-4880-b0c5-e25e5580c688 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.293552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.293552Z digest=sha256:a422bbd172a7e4960ade93f1a4b0ba5efe0ab820bbad2cb08703d45453991bb1

Observation 9aafee33-eba3-4100-aaed-5c0c03905092 · outbound

This paper cites Skywork open reasoner series, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Skywork open reasoner series, 2025

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:40.154755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:34.376071Z digest=sha256:0d171171610583609b9d1832e6bdb26444a21e1979aec3ee2ada2b0479924dae

Observation 1ea21ee6-5cfc-455c-84af-e57f4e4a53e8 · outbound

This paper cites Skywork Open Reasoner 1 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Skywork Open Reasoner 1 Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.446923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.446923Z digest=sha256:890a77bb6969812092b2cd79e4e1b9929a5f364744f0a4b941afc796b2435302

Observation 8ac4abad-be37-47b2-903f-749c58a81629 · outbound

This paper cites Measuring coding challenge competence with apps.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Measuring coding challenge competence with apps

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:40.048502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:34.541666Z digest=sha256:c23f140f63301102e20b70180b36dc2c423e6d1a597bd5143272823184fe34a4

Observation 125766b4-f36a-4f80-b7c4-f054e7b61b28 · outbound

This paper cites Measuring mathematical problem solving with the math dataset.Sort, 2(4):0–6, 2021.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Measuring mathematical problem solving with the math dataset.Sort, 2(4):0–6, 2021

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.839167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:34.631933Z digest=sha256:60d2375aeb0685a4eea8e54d4dbad431d65c05c58243ab5e76ac116f08cb4b3b

Observation 002ae11a-2cff-4bd7-98d6-c45ae10de270 · outbound

This paper cites OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.685646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.685646Z digest=sha256:2427c7ed7e2b0b22df418bcb7a37c917d080c19e68a6c49d9a6e5a519e03f5a6

Observation 5d3b1ec4-730b-4bef-b208-bb4a2e515337 · outbound

This paper cites Open r1: A fully open reproduction of deepseek-r1, January 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Open r1: A fully open reproduction of deepseek-r1, January 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.657016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:34.781301Z digest=sha256:29936ba406095d251a78a8b0c3d38ad89b7504652abaf0d515fca3a0c917115e

Observation c4d8b3cc-da20-4cf2-9e67-aad68a30c6f8 · outbound

This paper cites Qwen2.5-Coder Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen2.5-Coder Technical Report

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.858625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.858625Z digest=sha256:a4ac6775d63110c4e299627fcb8987b92514c49773e0273656c691923ed74ed4

Observation e0cb5a1d-9dde-429b-b25b-eedf95fd2419 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:34.924620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:34.924620Z digest=sha256:0b860688bd31a932db9a7d048d74b38b974bd2c494c5534563fe0587ade0aa8d

Observation 0f726602-e26b-45fd-964c-7a051a721913 · outbound

This paper cites Numinamath.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Numinamath

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.477517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:34.978623Z digest=sha256:4e945a5148d013e2fbec0247ef6e12cd37788ee2dd3349fa0fc8d53ebe1b833a

Observation 87136dd8-ece5-4de9-ab84-a61df9295c3b · outbound

This paper cites TACO: Topics in Algorithmic COde generation dataset.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy TACO: Topics in Algorithmic COde generation dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.095187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.095187Z digest=sha256:125351752b219efe03d2a6f9deb57a92bea9ec7142920e4db534b9edbab9befb

Observation 2ffce109-9866-43db-92d3-61e73d41ebcd · outbound

This paper cites DeepSeek-V3 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DeepSeek-V3 Technical Report

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.160615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.160615Z digest=sha256:f4b591522940918c63bc1a84b593ca123747d272327ebc8bac1082df7788d0b0

Observation 71f40547-feea-4652-8c01-086e7bd6adfa · outbound

This paper cites Is your code generated by chatGPT really correct? rigorous evaluation of large language models for code generation.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Is your code generated by chatGPT really correct? rigorous evaluation of large language models for code generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.332590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.244979Z digest=sha256:e4df8f553616a6ee5e9dac574de1fdb9003fdb3dec30a95b28e34c6a751f5958

Observation de862560-3c48-4892-8318-73a99d6e4565 · outbound

This paper cites Evaluating language models for efficient code generation.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Evaluating language models for efficient code generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:39.117026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.328729Z digest=sha256:5e7ca149fca80c6eb55ded10a975ae5a12faec4fe00936fc2a8beec377913216

Observation 9945ac94-99ee-4cd5-bc09-518bca2d4311 · outbound

This paper cites ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.394606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.394606Z digest=sha256:20a36e026221fec6a5852c52dd5f540043d84d5380def74a2f0771320924974f

Observation d6db3524-6140-4563-bca6-9ab0e975239c · outbound

This paper cites AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.455218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.455218Z digest=sha256:42f64ac897698a0d0604169a1c0ec4d94f0a5ed8030821bbdaf1308f756e919d

Observation 2afc52a7-56e5-438b-8c28-8ba909b50927 · outbound

This paper cites Deepcoder: A fully open-source 14b coder at o3-mini level, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Deepcoder: A fully open-source 14b coder at o3-mini level, 2025

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.888394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.528184Z digest=sha256:debf377e57100de79be3dfbcc57d7339e90d0a6a7ce29b6c8d27ea0f56e5f1cc

Observation ea4a2324-f54d-4376-a819-b33b50163490 · outbound

This paper cites Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.645287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.582507Z digest=sha256:bef604945d79444b84e8bdc0ee055398160644c77492d67789859e796710d433

Observation 2babe2d0-3c1a-4d25-a5a2-1d5a3562de6a · outbound

This paper cites AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.629509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.629509Z digest=sha256:6822b8ad7d56c1d0a770c340f4046afc99baf5ae261f0f97498316dcff091d69

Observation b79fcdd3-db65-41e7-8170-e304bc684a96 · outbound

This paper cites s1: Simple test-time scaling.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy s1: Simple test-time scaling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.685233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.685233Z digest=sha256:10464c3e0620f5b7d97de608112498582e58d9addfd6bf7452090abfd983e61e

Observation bc47100d-b5b8-42d6-b17f-5d18db604ff4 · outbound

This paper cites Learning to reason with LLMs, 2024.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Learning to reason with LLMs, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.405092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.755514Z digest=sha256:e0c13f9fe009c759b5b0e08209e7eb7fb86e32e3500cac3937f05436bc4ae631

Observation 1ab9a39b-735e-45a5-8a5c-45d69449b868 · outbound

This paper cites QwQ-32B: Embracing the Power of Reinforcement Learning, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy QwQ-32B: Embracing the Power of Reinforcement Learning, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:38.162217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.797155Z digest=sha256:7415d38b435d73d3f73fd7cc730a0ea8b39cb97322ef6ac91ee10156833e5c53

Observation c05608b2-c707-47d4-b036-421817536693 · outbound

This paper cites Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Areal: Ant reasoning rl.https://github.com/inclusionAI/AReaL, 2025

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:37.812618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:35.854088Z digest=sha256:f3b6c7bb65dd00705bc239ac4e96f4780d594e598fbf64da10590e092c388fb7

Observation 354f8ca7-2d52-4513-8a7d-67db3398f439 · outbound

This paper cites Seed-thinking-v1.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Seed-thinking-v1

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.912258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.912258Z digest=sha256:b11b104f256da789239bcf4d55dd6d5573e37f437b2efeb6225c245101dca413

Observation 2a1a5481-17f1-4988-bfad-ec0c516bf108 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.965118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.965118Z digest=sha256:f6e2258bcfac3d67b4886d48622b4527aad675cec1a9dd8f8e38c8ad3af3e26e

Observation f7566cfb-ade5-4ea6-b269-0b0c8818a398 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy HybridFlow: A Flexible and Efficient RLHF Framework

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.026976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.026976Z digest=sha256:46daa04a1eccf4f903349bc5b2017352df892f56a3972b01e1a6e78dc1c9278f

Observation 93e0b815-2d12-4c8e-a96b-0ec5d8764436 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:41:37.222993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:41:36.056695Z digest=sha256:94416910a50bf027daa5f88a0d433799761a347a2fe87db8e0372321f5f0be21

Observation 71a6e346-aac4-4dea-a418-2222643eb1ae · outbound

This paper cites Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.113617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.113617Z digest=sha256:872de084e2aead9a3172a19142bb55935f1733a609dbde501c14437905cd5612

Observation 56c7b4e5-03ba-44f0-8652-467e2c577f9e · outbound

This paper cites MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy MiMo: Unlocking the Reasoning Potential of Language Model -- From Pretraining to Posttraining

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.175903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.175903Z digest=sha256:09e5a4be5f4e929d611555690c9215625459a48f22f96d7a885279f2279e4efb

Observation 089a198a-d9b2-4258-9c7f-c451936bb0c1 · outbound

This paper cites Qwen2.5 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen2.5 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.237558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.237558Z digest=sha256:da72baa1785c40fbc5e2b99d0412eafd4885d2b83ef33777a25563ff82acd660

Observation a02be7d8-538d-4b6c-8f07-a8abbbed94cd · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.284404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.284404Z digest=sha256:728de672c60a30fda695d667fc80b025f711c7cc73b8e29433434d5b35db750f

Observation 653ecf0f-07fd-44be-b8f6-01027a8fe71f · outbound

This paper cites Qwen3 Technical Report.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy Qwen3 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.361667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.361667Z digest=sha256:b2b9e3627ad8ecab7baa5c5b564b417942929a8353e98691bf33026793bf33da

Observation b93c90ad-17e2-412f-aad4-b3c5f5b4155f · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.411140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.411140Z digest=sha256:ce4a292a1459637dcb470b1ac23b6e552cf7d95f6e5881902e6223557510d642

Observation 39ae18ee-d0a8-474b-9222-8cddd87ad256 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:36.479613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:36.479613Z digest=sha256:f092e190a405d2a3edeaab8f3ce47b0d2141bdfcf0cc6acfa806f025adb86a12

Pith citing papers

Observation e7e64ebd-a01c-41a6-b536-78b606eefee0 · inbound

DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation cites this paper.

DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Generation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T22:52:33.396332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:52:33.396332Z digest=sha256:43e67c50f0764fed59334e42862b0632f95cd0da2439415292de6b78ea64e51a

Observation 715a276c-f616-4c4e-a226-32d82389387f · inbound

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs cites this paper.

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:12.895530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:12.895530Z digest=sha256:0ab2e80463e516a7dd85ddd2930f74f7d3e726c0409cfff0e348c766d27e4fda

Observation 263d6ffd-74e9-4133-b666-486e59853516 · inbound

The Signal is in the Steps: Local Scoring for Reasoning Data Selection cites this paper.

The Signal is in the Steps: Local Scoring for Reasoning Data Selection AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:16:14.215851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T10:14:27.739531Z digest=sha256:53a76944380a058984bd06521100938d79f513a9170898daf7fd7c0190842cdb

Observation ac544d11-8b86-4391-9fb5-28d1cf1c98e2 · inbound

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation cites this paper.

Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-18T10:06:13.902007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T10:04:39.223895Z digest=sha256:04e4f1e43f39581118c7a79e970f65a548090185474dea427664a5fa137a27fe

Observation eead6158-8dfa-463a-a142-fcfe2b778c16 · inbound

Video Reasoning without Training cites this paper.

Video Reasoning without Training AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T09:12:08.416668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T09:12:08.416668Z digest=sha256:f5749a0630ad7d57b75f35a33196c4a9fc4d7cc3d92d1ed858bd4f5f05119020

Observation afb90df5-e7b0-44f9-8e70-6f30577c4aae · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:43:11.941650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T07:43:11.620446Z digest=sha256:436fa7bf7a638a20a08c98cf66e757a66ed6d71256501414e02cc50df6dbbec7

Observation d91f228b-8b9d-4d6c-8007-241b3ebc6d38 · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T07:31:49.069930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:31:49.069930Z digest=sha256:0d2a568510dddfa77e0054831b361f6a6e8d69d3fd91f61e59bcf35f9e5d4c9b

Observation 5d94fa74-f831-48c4-ab64-c5f8bb37ee83 · inbound

NVIDIA Nemotron 3: Efficient and Open Intelligence cites this paper.

NVIDIA Nemotron 3: Efficient and Open Intelligence AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 170

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T01:40:42.712572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-18T01:40:42.190369Z digest=sha256:fb3b99ec712f3d0f4190c66a4f1617b52053fa99074a14c73b6a712cd2a1bef8

Observation f0c6be84-0d9d-4402-ab99-02590d6f2afe · inbound

On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training cites this paper.

On the Non-decoupling of Supervised Fine-tuning and Reinforcement Learning in Post-training AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:48:00.530859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T14:47:52.094353Z digest=sha256:54491ad33836845b30bae3e07051e7c8b4539ceb7522f508aeafdc7763966b93

Observation 386d4239-5b69-4d4b-8170-b11d040d9dcf · inbound

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse cites this paper.

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:47:37.216769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T08:47:29.236561Z digest=sha256:a106d8f49ab69119fb86140c74f7bd48de866c18c143adfc47d8929f09d23c2b

Observation 1bfe6c99-0e3a-400c-b390-91593e889ac1 · inbound

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse cites this paper.

Attention Sink Forges Native MoE in Attention Layers: Sink-Aware Training to Address Head Collapse AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T05:50:24.502645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:50:24.502645Z digest=sha256:bae808d722e52267f28e0f677fbcfbaf194cd4be9fb9a6e3e8572ba643c0641c

Observation 1af9d3c6-ab40-4d68-a453-aac7e9427bf4 · inbound

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models cites this paper.

Entropy-Preserving Supervised Fine-Tuning via Adaptive Self-Distillation for Large Reasoning Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T05:28:13.898465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T05:28:13.898465Z digest=sha256:2b23336d350b92db56ab0c3cf9e7b538d8e2b6cf11101cc8f989cd6ce5d3153e

Observation 68335bce-2ae7-42b0-9ef4-251cf552bce8 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.537608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:f58f599790beb98d544f2c0662458ecc699ef803346d45483b41763f0d32803b

Observation 538c34aa-db4b-4a15-a4ed-22f28988bc3c · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.084430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:9226ec90e31389993dfd076d46baa206bacf2fb6e1925a09205e918fb6e764a1

Observation 6dabc6dd-63a6-434f-ab24-e836b93fdf50 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:06.791321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:06.791321Z digest=sha256:4d414a58a5861e4517472165c8a3a1d19ed11efb22bb1e3e81a9eb731dd6610d

Observation 9ab6c976-a761-443f-93a7-67ca89d3f58b · inbound

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data cites this paper.

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:18:21.791845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:18:03.185565Z digest=sha256:92fa0a5dc6acddf986317e98de1f363e24c13cf0202d079fca78717c9544b15e

Observation 84f59440-29af-4df6-917f-caa49c812b68 · inbound

Characterizing Model-Native Skills cites this paper.

Characterizing Model-Native Skills AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:56:11.635008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T05:42:49.694715Z digest=sha256:1f6127d082bf6534a3b8fd5d98277503e582327cb4a725941574e8e99307784c

Observation db8bbd0e-e3bf-45ce-af2f-bb7dc3e87169 · inbound

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models cites this paper.

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T02:45:57.822136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:44:36.697450Z digest=sha256:02f1b6024fe49eaad43471e45ca26846704fedee1400c5d804e457b208b61296

Observation f0020f30-01b6-496c-9681-05b896a5c9e5 · inbound

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models cites this paper.

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:03:50.645752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T23:01:17.957634Z digest=sha256:9d2f6c983c089004dbde063de2b36f9a2187e3b06b8927dd5c664a006a3a41a2

Observation aa0f8adc-f8b6-47f9-84d1-1f9495c39b3a · inbound

Post-Trained MoE Can Skip Half Experts via Self-Distillation cites this paper.

Post-Trained MoE Can Skip Half Experts via Self-Distillation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:03:15.213905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T12:00:35.496822Z digest=sha256:4c962f4611896369bfca3ee672ea277b309c1446b4f1550326cede1016b3f0d2

Observation 184f504f-b847-4fee-8101-a1d82d72c29e · inbound

Post-Trained MoE Can Skip Half Experts via Self-Distillation cites this paper.

Post-Trained MoE Can Skip Half Experts via Self-Distillation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:25:00.079365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T18:22:45.702572Z digest=sha256:7248d587a7cc77af88cf76a9568d9b51209abe3366912346fc75464f92a6b741

Observation 8630763e-b68f-41cb-944c-fdcacb880604 · inbound

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models cites this paper.

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.530928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T08:01:39.412431Z digest=sha256:9e843ed2b4fd8d892ab8bc1eac834d8eee2fb1adf6836de341109f1976948598

Observation e1dd7765-87b1-4ff2-9fbb-f236cda1a1cc · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 255

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:56:13.643184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:149445cc5d9d07759666cc052d3bf8b1df3f8d6ded497c15806ef22f3aac281f

Observation 5941e877-a830-468f-8e55-20c5f836ebe2 · inbound

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning cites this paper.

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T09:47:59.921475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T10:18:54.163862Z digest=sha256:9dbcb0539099d6564997785eed00e14cb8a67aa336802b5bbe6f08ac68fddb73

Observation 394020ad-4f17-45e5-b4a5-970d706b09fe · inbound

A Sovereign, Open-Source Foundation Model for German and English cites this paper.

A Sovereign, Open-Source Foundation Model for German and English AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-13T03:06:27.991558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T03:06:27.991558Z digest=sha256:848d4558ab62443c9bb4dccdec058be6f04a43e05be47964a238ba39b2da075d

Observation 9e440f94-c3f8-4a69-ad08-a6249f81933b · inbound

A Sovereign, Open-Source Foundation Model for German and English cites this paper.

A Sovereign, Open-Source Foundation Model for German and English AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T15:13:39.458378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T15:13:39.458378Z digest=sha256:5bf3d2a115c8e55df56422cce5947366315707aba0116e856f446653d8cfb61c

Observation 081a4e37-aac3-49e0-9ac7-af96fbcc5f6d · inbound

A Sovereign, Open-Source Foundation Model for German and English cites this paper.

A Sovereign, Open-Source Foundation Model for German and English AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T07:45:34.584748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:45:34.584748Z digest=sha256:b7c304cfde746cc65cd9521a9b32af798affb5932e8042432a829fc1d0a6217f