Pith. sign in

Paper Citation Record · LEDGER

Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2411.00750.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.00750 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:19.982278Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T01:48:51.021634Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 076a0e2d-1f97-4d53-bcc2-41fcb653ef49 · inbound

Self-Reasoning Language Models: Unfold Hidden Reasoning Chains with Few Reasoning Catalyst cites this paper.

Self-Reasoning Language Models: Unfold Hidden Reasoning Chains with Few Reasoning Catalyst Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:19.982278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:42:19.982278Z digest=sha256:aa2387073c668a9ce000bb90b031d10054ee0a4df49c19feabe75f40d0ac30a3

Observation 94980938-e7ef-45cd-afb0-5e77c8156dbc · inbound

Progressive Mastery: Customized Curriculum Learning with Guided Prompting for Mathematical Reasoning cites this paper.

Progressive Mastery: Customized Curriculum Learning with Guided Prompting for Mathematical Reasoning Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:54:07.719579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:54:07.719579Z digest=sha256:231f24c84a1cd649c9a40159343b57f57daa81dc97213e7f7e8235f3444356e7

Observation 9e9eb848-76b7-40e9-b642-e38be1eb0f71 · inbound

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training cites this paper.

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:48:51.024907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T01:46:21.744857Z digest=sha256:0b193a0b27c35bc3927e09d30eb510686cb319266d6cb3a9217d54e54aefc647

Observation 794f2014-a0ca-4860-9a9f-14bb5b8c85e0 · inbound

More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD) cites this paper.

More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD) Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

Reference 3

Resolution
malformed identifier
no resolver link, observed 2026-08-03T07:05:16.726478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:05:16.726478Z digest=sha256:3bea84d62081df21d2238d5234d5da828dffb9abc267a86d483c14788b800beb

Observation 850965d9-7ff9-4728-b12b-a4ce7b0d54ec · inbound

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs cites this paper.

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

Reference 100

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:45:59.445313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:58:10.013475Z digest=sha256:1b59e38c40cc595f05fec90911e569dd01cd4cc0e91ab7d90f46a00d5e6ce15a

Observation 905add15-a69b-41a9-870b-5c034c91271f · inbound

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning cites this paper.

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Mitigating Tail Narrowing in LLM Self-Improvement via Socratic-Guided Sampling

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:01:00.554757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:51:19.555272Z digest=sha256:7b7bfd402a504a0660477c0d533cab6b9368ea0822d3c239175d11d3c3d04159