Pith. sign in

Paper Citation Record · LEDGER

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning

As of 19 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2607.07690.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07690 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-09T02:29:12.366018Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact13
  • verified fuzzy1
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8b390d5b-bf7a-42a2-b48e-ce75570111dc · outbound

This paper cites DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning DeepMath-103K: A Large-Scale, Challenging, Decontaminated, and Verifiable Mathematical Dataset for Advancing Reasoning

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.868114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:b48a9cee33e332b5800a416a7b6c3f2382f4d41dd2ef2fdee327323e1aceea06

Observation f0f61310-7bfc-4e5a-98d5-f0c3f2e26e29 · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.844868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:dafc8dc485f41f40abd74eca2e75ebb5b406f882fc66d38505c805b6c0a88745

Observation 39e7199e-d52e-42b4-925c-420e2dcb72a1 · outbound

This paper cites Let's Verify Step by Step.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Let's Verify Step by Step

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.870304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:c33a87ef383506ab79529e2be8250813cb75cb43bce7341ef288b2054add634f

Observation 6e9aeb06-992a-44be-a519-1256a02baf2c · outbound

This paper cites Mixture-of-Agents Enhances Large Language Model Capabilities.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Mixture-of-Agents Enhances Large Language Model Capabilities

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.865745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:39960a59be1fd36428afa2a5714f76af726a1b2c4c419c7ce74614cb5a4e69da

Observation f4478556-1642-4af0-ad24-9a002f7db2fc · outbound

This paper cites Rethinking Mixture-of-Agents: Is Mixing Different Large Language Models Beneficial?.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Rethinking Mixture-of-Agents: Is Mixing Different Large Language Models Beneficial?

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.863433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:bd11b7fe966c374ce79109e383c17ea39c7e983fdaec3194ef7b5cef1e641874

Observation 8b4461bc-13ee-49a2-89a6-2bda15b70083 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.856386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:b5e1d1c4b24cb569fbe6c318c2c16ea2db3d7c94a175cb6ae4a25ee139fa7776

Observation e4a74279-336c-4cc9-b890-8092b9490709 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.858732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:00052d786aa0cb05985ec4221021387c2575de193a75e8bfe9cf572efe29db06

Observation 8c4eb0d9-fd91-4852-929b-5193a0ddc3e7 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.872758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:af7ba7b9b8846cf83c5c5b58084b5c7bf744d529eb18ef117c2f6b91365063fc

Observation 3dc2e27f-1376-472d-9505-122daa81c3d4 · outbound

This paper cites R-Zero: Self-Evolving Reasoning LLM from Zero Data.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning R-Zero: Self-Evolving Reasoning LLM from Zero Data

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.849536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:b795cd8f4c6485dec95573e95e5751d0d98418e8d9a30e1b9364cba97490b387

Observation 35730966-7e03-482c-b0b2-b0431a46b127 · outbound

This paper cites Improving Factuality and Reasoning in Language Models through Multiagent Debate.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Improving Factuality and Reasoning in Language Models through Multiagent Debate

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.854070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:8a4d5fcdc60b0966ff74fdc3a798615e0d783eaf394c1e2222da8ef68bb95710

Observation 247014f7-214e-47ce-bcf1-df8bbdd196f4 · outbound

This paper cites AI safety via debate.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning AI safety via debate

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.861079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:9d5f0687ca431c0f4cc15b0519111014772eee01c302a98f4094c3a2fb20b032

Observation 4b1b1f36-2229-4363-8bad-ee62b3ab3fd7 · outbound

This paper cites Prover-Verifier Games improve legibility of LLM outputs.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Prover-Verifier Games improve legibility of LLM outputs

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.847137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:62defa8c09d4afd60866ae9b1a3f064ccbdf18a3ad13b9d52de53a68b598d697

Observation 43195baa-e0c0-4892-a195-0b1b34dce7b1 · outbound

This paper cites The 2026 AI Index Report, Chapter 2: Technical Performance.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning The 2026 AI Index Report, Chapter 2: Technical Performance

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T02:35:54.386630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:b513eedbb67c64f1f093e5d9fda8f32d0914cb6f38d3de4280ee2ec8f6245678

Observation 059be145-2ae5-4663-88c4-898bbe5c21a2 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-09T02:35:53.851785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-09T02:29:12.366018Z digest=sha256:f1067eeaa8b52c9a448d90e5c4e3316942182f097a59acd52ff9cb106027f933

Pith citing papers

No inbound Pith citation observations are available.