Pith. sign in

Paper Citation Record · LEDGER

GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2310.12397.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.12397 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:12:13.428065Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

11
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c30d184e-8fd5-44b6-89be-c49c8358f46d · inbound

Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning cites this paper.

Veracity Bias and Beyond: Uncovering LLMs' Hidden Beliefs in Problem-Solving Reasoning GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:12:13.428065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:12:13.428065Z digest=sha256:fb4159bcffa62420aa4316c84a472518300c67d1118af7ec30fef88a7eb38d01

Observation 0d1cd1f9-3767-42b5-80d3-572571960587 · inbound

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models cites this paper.

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:00.039133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:00.039133Z digest=sha256:c1fab6c7ebf18a12913d1faf1ef010d296a63ee6380e930ca1228b6d9a31f653

Observation e490b109-c17f-4cd5-a8c2-c1fc2c6ca889 · inbound

It's Not That Simple. An Analysis of Simple Test-Time Scaling cites this paper.

It's Not That Simple. An Analysis of Simple Test-Time Scaling GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:05.621221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:05.621221Z digest=sha256:77dfac1c31086d0f2151a11ebeebffe1b9a3693d41b3df8a1fcdd7b6f6ef7120

Observation c2544d17-6930-4b8d-88d5-f83cfb1cdac9 · inbound

Lightweight Language Models are Prone to Reasoning Errors for Complex Computational Phenotyping Tasks cites this paper.

Lightweight Language Models are Prone to Reasoning Errors for Complex Computational Phenotyping Tasks GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:06:07.559969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:06:07.559969Z digest=sha256:41d4a3c1dbeb53be8ee24ce7d4ccaa1467c028ba971830c69025b3d73e0d7659

Observation 56886bbe-0b59-4a90-8aae-f3d38fa406a8 · inbound

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces cites this paper.

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.605389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T02:57:15.521594Z digest=sha256:d2f57bfcdd3a5e9f6f2f8994e0cf9925b25e690cf81a7a281df67947e9823b93

Observation cf2eba2c-f2d4-4fdf-897d-fe6536776d99 · inbound

Roll Out and Roll Back: Diffusion LLMs are Their Own Efficiency Teachers cites this paper.

Roll Out and Roll Back: Diffusion LLMs are Their Own Efficiency Teachers GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:42:46.111972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T20:41:48.063871Z digest=sha256:ced4b1a283b7708a8f77d51c91449b9bacc1cc85dd4ebc0cad5859f61473f47c

Observation 40139c4b-32fe-4dde-99d6-4ca72fa8a5ae · inbound

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning cites this paper.

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.375155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:23:24.536258Z digest=sha256:dccad8db473c4a599d57149f30017c267cea9dc1e673c233cbff115893fed7f6

Observation 06a3d703-496c-4a31-aaef-27d37cbbcfb3 · inbound

The Self-Correction Illusion: Role Relabeling Gates Explicit Error Flagging in Large Language Models cites this paper.

The Self-Correction Illusion: Role Relabeling Gates Explicit Error Flagging in Large Language Models GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T01:31:29.251945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T01:25:07.890796Z digest=sha256:fa08601af2731d92af0776a608691024373f628767b31a809c29419d5d429743

Observation 3384582e-a352-49d0-a3ef-65f30119565d · inbound

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination cites this paper.

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:07:23.497033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T19:58:32.016341Z digest=sha256:1d1e42d80d6545db7b202e058b87331fa50ec86ef9500d819f99dda59ae13cf4