Pith. sign in

Paper Citation Record · LEDGER

Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2308.07921.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.07921 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:33:53.166179Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:05:47.722524Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f045e8db-7a38-4741-ba4c-2e149467e263 · inbound

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning cites this paper.

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-17T23:46:39.679051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-17T23:46:39.330438Z digest=sha256:534aba4294c2f793afa1215b8e594913a1dd06081c0952621fbbce21d58f4926

Observation 347f1a27-5ce3-441a-bc87-350987c6a856 · inbound

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving cites this paper.

ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:19:36.202672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-19T09:19:35.918151Z digest=sha256:3af38bf921adc4bf9c008fee660c7dd9010f9a21b165f782bcf670739e1962a3

Observation a2fa2715-fbc3-4580-9f60-aa9bdc76923a · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:27.558572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:c7c515e09fb3137fe36732c9e2a8083012cf8006164c8d88a8b7fd9bb46b1780

Observation 2c6470d1-797e-4621-bdf4-c35bfb147b42 · inbound

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models cites this paper.

SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:03:26.818551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T03:03:26.723464Z digest=sha256:49a089e589ac1fee6d3dc60b6ca017fb8f4a1c5bf0c7c34327c68f01fe85c82b

Observation e7facf7c-f0db-41e4-b826-a2b3d64a6350 · inbound

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? cites this paper.

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:29:30.194704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-17T01:29:30.032408Z digest=sha256:c4d0138acbaf8f680dc96c2d59c46229247e6b183897ff42ef33f2fba9db9738

Observation f0975690-70c8-4c1d-991c-b99feeaa2ca5 · inbound

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark cites this paper.

Polymath: A Challenging Multi-modal Mathematical Reasoning Benchmark Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:05:47.726018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-23T20:03:38.336841Z digest=sha256:a2a5630f0985412a0368a56787b78153073d61cf9a4b50fb4af2f6e9dd29aa46

Observation 0d5e8ca4-d28e-4697-85e7-a1103940ea43 · inbound

Curriculum Demonstration Selection for In-Context Learning cites this paper.

Curriculum Demonstration Selection for In-Context Learning Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T11:33:53.166179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:33:53.166179Z digest=sha256:f36ec2bdfb95f091466a636dff0d12dc617e8b9b2f723b52331a839ccfb0e80b

Observation ec6e52b0-60f0-4bfc-a642-8abeb6402a57 · inbound

Mars-PO: Multi-Agent Reasoning System Preference Optimization cites this paper.

Mars-PO: Multi-Agent Reasoning System Preference Optimization Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T10:39:36.971073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:39:36.971073Z digest=sha256:62992a92ed85855e38eaab2597448e50a2cc183be520ffec0eff5827553829d9

Observation 74e89f74-884e-4dff-b20c-b78bc5125263 · inbound

Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents cites this paper.

Improving Physics Reasoning in Large Language Models Using Mixture of Refinement Agents Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T05:03:21.042575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:03:21.042575Z digest=sha256:8bb9229e8e05540d588b04bc38cedab6289a70406998c3d026ddc5003a2644e7

Observation be4865e5-94dd-4457-86ae-ec186449e587 · inbound

RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models cites this paper.

RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T23:08:16.618728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T23:08:16.618728Z digest=sha256:cdb9c1713c66e2204f44415ecbad8dd2551a49fc9f03545ea9a5d7423dc2ffa9

Observation d85234bd-00c6-49a1-8ab6-478898fcdea1 · inbound

Zero-Shot Verification-guided Chain of Thoughts cites this paper.

Zero-Shot Verification-guided Chain of Thoughts Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T17:51:25.472902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:51:25.472902Z digest=sha256:9295333d6a6fc2eca23dc718a53f6d886cdb92f935705e75addb2fb14012af29

Observation 23df265d-f769-483c-a362-c3dfa64b58b4 · inbound

UGMathBench: A Diverse and Dynamic Benchmark for Undergraduate-Level Mathematical Reasoning with Large Language Models cites this paper.

UGMathBench: A Diverse and Dynamic Benchmark for Undergraduate-Level Mathematical Reasoning with Large Language Models Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-10T15:40:39.926795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:40:39.926795Z digest=sha256:ad8e042e8bebde6268315f1df5dadae29635adba3a5f3ddc196624749a1a5552

Observation 4d0ba911-8191-4bfc-a10d-c463fe9d0f0e · inbound

CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance cites this paper.

CodeSteer: Symbolic-Augmented Language Models via Code/Text Guidance Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T12:16:27.449334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T12:16:27.449334Z digest=sha256:ce74e0f98f7b6f279fa12515977cd9094f74cc3f5ef0564b9e3fc9b8587d0451

Observation 6ab55779-bb30-40fb-b498-6333e404cc1b · inbound

Large Language Models for Predictive Analysis: How Far Are They? cites this paper.

Large Language Models for Predictive Analysis: How Far Are They? Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:26.433933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:05:26.433933Z digest=sha256:7855f9920fa793898df8ec07d63ef94632a0713d5d258c4348108253d8702e8b

Observation 0137b9ff-af69-4a7d-b497-3944549369f4 · inbound

Synthesis by Design: Controlled Data Generation via Structural Guidance cites this paper.

Synthesis by Design: Controlled Data Generation via Structural Guidance Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:34:23.271425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:34:23.271425Z digest=sha256:e4d741dc7e38dafe8be663e5452d52a5442f275327e334dc1ae047bdb62d0848

Observation 75d6ff91-e960-4f8e-af03-926e0ab8e6dc · inbound

Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning cites this paper.

Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T00:34:39.345662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:34:39.345662Z digest=sha256:b27658429278790c675e0155d349ab27b8ca1eb4635223fa4fde9f3723a84ecb

Observation ebb0946f-dda3-43fa-ad22-8ee8241310c6 · inbound

CodeAgents: A Token-Efficient Framework for Codified Multi-Agent Reasoning in LLMs cites this paper.

CodeAgents: A Token-Efficient Framework for Codified Multi-Agent Reasoning in LLMs Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:19:20.707339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:19:20.707339Z digest=sha256:62041542fa64da60031ad14007b532089a4c05be763806cabb668904f2f0a17a

Observation 4fc9fc3a-9ae6-4abc-bdbe-ebe7479750bb · inbound

MENTOR: Reinforcement Learning via Flexible Teacher-Optimized Rewards for Tool-Use Distillation cites this paper.

MENTOR: Reinforcement Learning via Flexible Teacher-Optimized Rewards for Tool-Use Distillation Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T08:55:58.819366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:55:58.819366Z digest=sha256:f714bb3d2dd1258e3af931cdea0e4bc062fad99321d2ef879e87e7c5cef85b44

Observation 4195cfc8-f499-430d-9dda-4b8b96997c42 · inbound

NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise cites this paper.

NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:56:07.223575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-08T16:54:17.663989Z digest=sha256:82ee388980e4c67fe5d75e38c7689b29620296c03ed1c6b11b86e6005b9345c8

Observation cdacb0a2-0273-4001-bb04-df4c0a37a362 · inbound

Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation cites this paper.

Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation Solving Challenging Math Word Problems Using GPT-4 Code Interpreter with Code-based Self-Verification

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T15:30:23.485228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T15:30:23.485228Z digest=sha256:befef0b9808c535ed9c1bb239fa3c68f618f963f1e48b343c0a9c167fe6ba28f