Pith. sign in

Paper Citation Record · LEDGER

Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2411.16579.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16579 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T11:11:12.780313Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.141555Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7aacda37-b046-471e-be59-73e724a8bec2 · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 171

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:23:14.944730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:c04e5c35f0802c6c329ee21df8f9b1666b820a925898172221c00a7925eee7d7

Observation 3396766a-ad2a-4185-b315-f00feb8a3e76 · inbound

Learning from Natural Language Feedback for Personalized Question Answering cites this paper.

Learning from Natural Language Feedback for Personalized Question Answering Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:56:53.045875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T22:55:16.228037Z digest=sha256:30d430d4db1100a2df868bbb8d7986cabdde0db478413748daa91a066cc48fa4

Observation 6eafc153-32dd-404d-9036-6b9ae1198fbe · inbound

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation cites this paper.

XRPO: Pushing the limits of GRPO with Targeted Exploration and Exploitation Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 1992

Resolution
unresolved
no resolver link, observed 2026-08-04T11:11:12.780313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:11:12.780313Z digest=sha256:8c6383fd8e32f02e3e2f3226f63c0c18ad4e948bb7134df41d69ac61f3686672

Observation f5360fa2-2d32-4f45-aa09-c34df55eb6aa · inbound

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning cites this paper.

No More Stale Feedback: Co-Evolving Critics for Open-World Agent Learning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:03:04.330035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T16:01:48.789986Z digest=sha256:7125bc8628eff2abc4e735250d74795788eaa500187472dcb305b4dcccaf5ade

Observation 6207f2ce-2e63-428e-8042-23e0caf6e0ae · inbound

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning cites this paper.

STRIDE: Learnable Stepwise Language Feedback for LLM Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:49:00.673885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T20:47:16.236629Z digest=sha256:ba0285abea26fd4c71e0f87ba8aeae0d19d8de91de6f106b7d5878c3f8c12028

Observation 39681fe5-9020-4b3e-8b23-27317878f03d · inbound

A History-Aware Visually Grounded Critic for Computer Use Agents cites this paper.

A History-Aware Visually Grounded Critic for Computer Use Agents Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:17:40.143173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T13:20:32.432002Z digest=sha256:34ea54032bf102ef7190ef27dce1494709c15649c9129ffd656e22e325488003