Pith. sign in

Paper Citation Record · LEDGER

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback

As of 8 August 2026, this Paper Citation Record lists 7 of 7 outbound references and 0 inbound Pith citation observations for arXiv:2605.28010.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.28010 v1

Coverage vector

measured 7 of 7 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:10:43.791548Z

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

7 of 7 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch6

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 303557f4-bbf9-4ae9-822e-0213d69acacc · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Training Verifiers to Solve Math Word Problems

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.642494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:67a1daa14f63cc0ce1335afa0a6278cde70a8081e7c49a268b5f23abe18aa6f3

Observation f1a743ca-44a0-4527-8789-57adf9d71484 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.644921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:9e75284a3b977d72430133e4f3ea0cfecc5a89299ac5852678c41ffed2a93fd4

Observation 820926d8-3a0d-43c6-aef3-1709b9d356fa · outbound

This paper cites OpenAI o1 System Card.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback OpenAI o1 System Card

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.635435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:726f6d5c9ef8a36435836fd3e722634b09f837d1f3b9b8d11b0331c35a3725fa

Observation b5645702-7a32-41bb-8a32-7793d0360542 · outbound

This paper cites Spice: Self-play in corpus environments improves reasoning.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Spice: Self-play in corpus environments improves reasoning

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:13:26.640194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:55d9ec3aa2a06ee260df945923f263c085a84f86126aa47d4f0ca6712415b744

Observation 7af0f012-17b3-4b2e-af7e-23559914b313 · outbound

This paper cites InAdvances in Neural Information Processing Systems (NeurIPS), volume 37.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback InAdvances in Neural Information Processing Systems (NeurIPS), volume 37

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T12:10:43.791548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:e5b1e38b062b70bf51e1e6eb5690f5b7ffed7aab0779b01a926ddbc3e46c30f2

Observation f94c1dde-40e0-4185-8ee6-f222e5c86b01 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Proximal Policy Optimization Algorithms

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:13:26.647159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:87c1fcafac987fcb05f9b57a26a673a9db09b9ff4c8c8178da7728a54d066091

Observation 63f226a2-7a55-4e16-b448-2faf98a44964 · outbound

This paper cites Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement.

Confidence-Orchestrated Self-Evolution against Uncertain LLM Feedback Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:13:26.637785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:10:43.791548Z digest=sha256:a835f657dac32c43c0f2619485b92c6c5c867c5aab69e5a593504e0849ca11f6

Pith citing papers

No inbound Pith citation observations are available.