Pith. sign in

Paper Citation Record · LEDGER

GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2402.10963.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.10963 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T19:39:35.508973Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:47:27.997707Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c228b56a-1f82-4db3-be15-1a584089f83e · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 131

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.662648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:aecc2b6526eb2b47a9d9cb0ddf79476e0d85c25bccd86459b88c4768f7b65530

Observation 3968d1a6-eb47-4860-8b0f-ed7a825bba9b · inbound

ViRAC: A Vision-Reasoning Agent Head Movement Control Framework in Arbitrary Virtual Environments cites this paper.

ViRAC: A Vision-Reasoning Agent Head Movement Control Framework in Arbitrary Virtual Environments GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T19:39:35.508973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T19:39:35.508973Z digest=sha256:0c237625941e17fafb9b7d06a86eee44dd95a7cd0966d2a0b23cb8162e98cb9e

Observation 58f50a08-3af2-4c3f-a900-dd231a7ac5dd · inbound

Decomposing Elements of Problem Solving: What "Math" Does RL Teach? cites this paper.

Decomposing Elements of Problem Solving: What "Math" Does RL Teach? GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:05:28.672657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:05:28.672657Z digest=sha256:34b621f4d81f7b3981af06547cbd3abea3ad30055fb8b4c64527e60a23f021de

Observation b210253c-344a-4e54-93cf-2142dcf91728 · inbound

Boosting LLM Reasoning via Spontaneous Self-Correction cites this paper.

Boosting LLM Reasoning via Spontaneous Self-Correction GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:30.670012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:30.670012Z digest=sha256:46b0c5285a9d1a6897667a85a27487322c106f349b41cd10d88fe68a07e865ff

Observation b44131e8-95ac-4c9b-991f-64554fd390d4 · inbound

AnnoDPO: Protein Functional Annotation Learning with Direct Preference Optimization cites this paper.

AnnoDPO: Protein Functional Annotation Learning with Direct Preference Optimization GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:47:20.679825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:47:20.679825Z digest=sha256:c32c6b96feabf607e9e80d61f473923fd91200fd5690785cd439af54118df59b

Observation dd6a5a69-1ef3-41bc-ac74-8854b124ef22 · inbound

CORE-KG: An LLM-Driven Knowledge Graph Construction Framework for Human Smuggling Networks cites this paper.

CORE-KG: An LLM-Driven Knowledge Graph Construction Framework for Human Smuggling Networks GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:40:57.287407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:40:57.287407Z digest=sha256:c1fb0fbc7ef7016358d5ca9fa1455b0124013ebaa55a1eae8033935571b9095e

Observation 91bf25c5-f553-482d-bf85-9d2a3fe134af · inbound

Hallucination Detection with Small Language Models cites this paper.

Hallucination Detection with Small Language Models GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:11:43.320414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:11:43.320414Z digest=sha256:e83d858d7ffcb0b87447c66701d8109a08d7de4f42cd42a7e45dca1823a38bdb

Observation 9f34cb8c-8df1-4872-9ef5-ba94cc628c77 · inbound

Data Diversification Methods In Alignment Enhance Math Performance In LLMs cites this paper.

Data Diversification Methods In Alignment Enhance Math Performance In LLMs GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:26.369435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:43:26.369435Z digest=sha256:de636d5a36f94df727cdcdbf12e0bcf4f1a6ca92244d8322e40e94174396e460

Observation 63882c2d-434b-4b7d-ae1f-80fc475e7235 · inbound

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning cites this paper.

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T22:58:34.400113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:58:34.400113Z digest=sha256:9ba2b74bd978d811e1aef472ad00f8fc3796bb41042f700c84be929d452f0b80

Observation 4c8d487d-2010-4224-a7bd-2a4cb2657a84 · inbound

Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions? cites this paper.

Inverse IFEval: Can LLMs Unlearn Stubborn Training Conventions to Follow Real Instructions? GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:16:51.124811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:16:51.124811Z digest=sha256:2733161f80f1c8fe2e1a463e6394ff17aa038be4d4f887d0cea206eac04bf1f6

Observation 32443d28-6b63-4d5f-8aa2-5226406af71d · inbound

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation cites this paper.

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T03:04:43.666446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:04:43.666446Z digest=sha256:98e2cb9a183d73b8149ecffdc7cbc0c739640e65b7038b48c08f4cca9e414123

Observation 768d9311-f46e-4f9a-9b47-0f619d380ba6 · inbound

CATPO: Critique-Augmented Tree Policy Optimization cites this paper.

CATPO: Critique-Augmented Tree Policy Optimization GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:47:27.999096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T19:30:07.971424Z digest=sha256:d7aa0b8bffc4ffe6bf9bb0fbbc020f1ce3b149990b3286d12be2f31c32757620