Pith. sign in

Paper Citation Record · LEDGER

REFINER: Reasoning Feedback on Intermediate Representations

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2304.01904.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.01904 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:17:57.094974Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

31
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e3096acb-1c86-440d-8185-3dbc94d3d801 · inbound

Reflexion: Language Agents with Verbal Reinforcement Learning cites this paper.

Reflexion: Language Agents with Verbal Reinforcement Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:51:41.960105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T13:51:41.915864Z digest=sha256:84c46859916ea01b3a70a83dcfb527f432f1c4552b30cf988109df3a9dd76777

Observation 9f0aa8a9-d8e8-4412-8f4f-b1b3b96b38ec · inbound

Reasoning with Language Model is Planning with World Model cites this paper.

Reasoning with Language Model is Planning with World Model REFINER: Reasoning Feedback on Intermediate Representations

Reference 95

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:49:29.040924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T01:49:28.796581Z digest=sha256:6e216c2dc674c84bd33c9c12f6e1d2f217e24548f07329ebe7317554acbd7533

Observation 470f6a98-4ba4-4c7e-a9f0-83603fff0859 · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:26.680353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:a47718a908385ded031fcd48569eec2ce3f85565ff6e28edfd25678e0e3589d2

Observation 5ec55e9a-2837-4dd8-8f29-f699904138f3 · inbound

Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection cites this paper.

Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection REFINER: Reasoning Feedback on Intermediate Representations

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-12T14:15:11.188259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-12T14:15:10.907921Z digest=sha256:cc9e42d98db364dbbb4478dc286d22e621cb4feba677d99355001a6a8dc7adc6

Observation 32e04081-216d-4744-b80e-e77d1c1c946e · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning REFINER: Reasoning Feedback on Intermediate Representations

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:04:10.395880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:e435086c077cee38d466332494ed59030c0a08bc40c5c8f5c5582a8f27cfe45a

Observation cd4d6fda-10f7-4d9c-bab5-dbb91c7f2cad · inbound

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods cites this paper.

LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods REFINER: Reasoning Feedback on Intermediate Representations

Reference 182

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:11:13.174753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T23:08:34.312466Z digest=sha256:198b7aaabc68f2baf3b0e79bbeececdb990acbeb57d3e232a7e721333730f245

Observation 787420ba-b6d5-439a-9c92-1643ad5b01fe · inbound

Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression cites this paper.

Can Compressed LLMs Truly Act? An Empirical Evaluation of Agentic Capabilities in LLM Compression REFINER: Reasoning Feedback on Intermediate Representations

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:57.094974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:57.094974Z digest=sha256:2752eb6d2949eb5216a5beea63eb5592644408bd7ed581f38207db7d5d80936e

Observation f155d4f3-4747-4e7c-872a-8db3d2a0e2ea · inbound

Boosting LLM Reasoning via Spontaneous Self-Correction cites this paper.

Boosting LLM Reasoning via Spontaneous Self-Correction REFINER: Reasoning Feedback on Intermediate Representations

Reference 1994

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:30.704566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:30.704566Z digest=sha256:28d6ec7ce800ed986915161e1d8beac479ee87b42e7fa7bfcae6d5632c013c88

Observation e8ed16b5-574f-449e-ac8c-df0a68763491 · inbound

Grammar-Guided Evolutionary Search for Discrete Prompt Optimisation cites this paper.

Grammar-Guided Evolutionary Search for Discrete Prompt Optimisation REFINER: Reasoning Feedback on Intermediate Representations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:39:41.448247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:39:41.448247Z digest=sha256:250fb28a679bbf0e926829a6fbefe4def3cd5e94602eec2e55ad48005f8ff515

Observation 839d50b0-0a46-466d-8eba-e3b81bc01d02 · inbound

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems cites this paper.

R4ec: A Reasoning, Reflection, and Refinement Framework for Recommendation Systems REFINER: Reasoning Feedback on Intermediate Representations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T14:58:26.461427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:58:26.461427Z digest=sha256:59e708140b5d423a35354f1faa724c0a488c1cc88e48be73ee8a576c84f68361

Observation 9487328a-74a0-4e0d-aa13-1c9f5f3ca051 · inbound

CS-Agent: LLM-based Community Search via Dual-agent Collaboration cites this paper.

CS-Agent: LLM-based Community Search via Dual-agent Collaboration REFINER: Reasoning Feedback on Intermediate Representations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-05T21:02:46.430138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:02:46.430138Z digest=sha256:ecb1befcfcaccbc85d200087adbd9a497b34c2a1a3063b60745e427075329177

Observation 8a44883e-0ada-4820-84d8-c33b2134b237 · inbound

User-Assistant Bias in LLMs cites this paper.

User-Assistant Bias in LLMs REFINER: Reasoning Feedback on Intermediate Representations

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:51:53.124790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T22:51:02.400926Z digest=sha256:fd408a36d854aeb2644092b608ff706cd1d61ab7dcb5993d0f215c511bc75567

Observation 070411a4-4247-4c6e-8087-67a3321a8953 · inbound

Context Learning for Multi-Agent Discussion cites this paper.

Context Learning for Multi-Agent Discussion REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:10:45.503755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T08:08:39.182921Z digest=sha256:650d2a15e1102d10d7c7b931d4d8226e7f92e520fb510ec57fe2e29a2e4890cd

Observation 63d3dc15-c6aa-404c-a389-34831506a622 · inbound

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection cites this paper.

From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflection REFINER: Reasoning Feedback on Intermediate Representations

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:30:51.458271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:04:32.931596Z digest=sha256:d88ebce67761174a3a8680a4743a4ade6dbe5fb07de41eca800a154a32d196f8

Observation 9e949e25-3d58-4b13-9b6f-57207a600d89 · inbound

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping cites this paper.

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping REFINER: Reasoning Feedback on Intermediate Representations

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:29:47.265335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T00:26:45.372232Z digest=sha256:4efc1e689ecc23fa264018a5145b300abc23a304ad779af0f00cb8d2724da1ff

Observation b1e15fc0-56ec-4a04-9e12-7452d6c454fd · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding REFINER: Reasoning Feedback on Intermediate Representations

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:16:05.856887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:0a26e5b7f79603f5cfc7a337557bc66f91b6e440aa842c062fd28d92395015fb

Observation 34da90e0-a87a-4878-a855-950eecb07ac2 · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination REFINER: Reasoning Feedback on Intermediate Representations

Reference 102

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.906282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:bf3a49c7f5ef6f07d75e50d636ba02ae53ee3abe5f7e82b411a014f108ed6d1f

Observation f6518443-64ab-4281-bbf2-e143d4ed1659 · inbound

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation cites this paper.

OPD-Evolver: Cultivating Holistic Agent Evolver via On-Policy Distillation REFINER: Reasoning Feedback on Intermediate Representations

Reference 117

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:48:56.458732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T01:07:49.603969Z digest=sha256:3d53f0bf5186c4892d9f35076fc92cad21d789e4a856a052fcb930e7d122afdd

Observation 395fcd03-678b-4aae-b706-19ff420863c6 · inbound

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers cites this paper.

VTOS: Learning to Orchestrate Vision Tools by Co-Searching Solutions and Observers REFINER: Reasoning Feedback on Intermediate Representations

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T21:40:08.308378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T21:36:17.228002Z digest=sha256:5f4b4f8bbe5b32a28e712bf2227af685a31ca166ce5fe89df8349e2e8a1feb57