Pith. sign in

Paper Citation Record · LEDGER

Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2308.07921.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.07921 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T15:30:23.485228Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:05:47.722524Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f045e8db-7a38-4741-ba4c-2e149467e263 · inbound

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning cites this paper.

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-17T23:46:39.679051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T23:46:39.330438Z digest=sha256:386a0ecd6f1a7ea264d5fdbf93d67c2b034197d2a154b92b8d44bb042a3866a2

Observation 347f1a27-5ce3-441a-bc87-350987c6a856 · inbound

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving cites this paper.

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:19:36.202672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T09:19:35.918151Z digest=sha256:04678aaae1f33460e0d2b9a31a87cabea9589533e14a34bd74689677c5a15c45

Observation a2fa2715-fbc3-4580-9f60-aa9bdc76923a · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:27.558572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:29d04db0c49bb1ee201f5ef201f6bb5de5e381db392d94724fd841c162ae8a05

Observation 2c6470d1-797e-4621-bdf4-c35bfb147b42 · inbound

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models cites this paper.

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:03:26.818551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T03:03:26.723464Z digest=sha256:f8bb685e8ef393d40493f29006781c64a0d8571ad9633e0ad524b2b57abe900f

Observation e7facf7c-f0db-41e4-b826-a2b3d64a6350 · inbound

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? cites this paper.

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:29:30.194704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T01:29:30.032408Z digest=sha256:367b07609a4c1257b5ff1bc38a448a8374c8a5c3c216afe58eb3222b50bb4ee3

Observation f0975690-70c8-4c1d-991c-b99feeaa2ca5 · inbound

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark cites this paper.

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:05:47.726018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T20:03:38.336841Z digest=sha256:1a45bd650d42b55f73a9ddbce90fa717dded211e3ddbbdc9c9ab1f5108096084

Observation 4195cfc8-f499-430d-9dda-4b8b96997c42 · inbound

NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise cites this paper.

NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:56:07.223575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T16:54:17.663989Z digest=sha256:132d9bebfb8fea19b24f47ec4a6703ba8f9ad61a52dbaac83dd53a89a7d677d9

Observation cdacb0a2-0273-4001-bb04-df4c0a37a362 · inbound

Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation cites this paper.

Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T15:30:23.485228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T15:30:23.485228Z digest=sha256:39d4467f0bfe5a3ede6a50bf2bc4db8c033a9aa48170fde0f248ec1680da36ba