Pith. sign in

Paper Citation Record · LEDGER

Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2401.09395.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.09395 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:29:23.349868Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bfcd8e04-2e58-4641-900e-b13f5109db61 · inbound

Hallucination is Inevitable: An Innate Limitation of Large Language Models cites this paper.

Hallucination is Inevitable: An Innate Limitation of Large Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:38:43.493535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T20:38:43.411206Z digest=sha256:773a02b43aa2ba318b01b0eefda8174e8b5d271caf36e3d1698d33d78901dc09

Observation b9bdcc29-efe5-424a-abc7-a9a3de7afc99 · inbound

GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models cites this paper.

GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:42:12.068947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-15T00:42:11.891829Z digest=sha256:6009bb4b0af662eecdf55accf6e120ccb186714c593344ba685cb65527c84f4c

Observation fb0c0006-fe86-4b6e-8b89-1913a40b2344 · inbound

In Context Learning and Reasoning for Symbolic Regression with Large Language Models cites this paper.

In Context Learning and Reasoning for Symbolic Regression with Large Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:53:21.215668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T18:50:40.378720Z digest=sha256:3b83572274ba73e963d7ca68b963c88ec38a805727b5b5be3529bf0f28509167

Observation cf364608-c4ac-4868-a8dc-109150435dba · inbound

Evaluating the Robustness of Analogical Reasoning in Large Language Models cites this paper.

Evaluating the Robustness of Analogical Reasoning in Large Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T15:29:23.349868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:29:23.349868Z digest=sha256:14b2f030bf7ea629906a84c036450eb0a4e43689d5ee661942ac2c5c7386c654

Observation e70a864e-b1de-4e98-b304-359b3fa79628 · inbound

Understanding and Benchmarking Artificial Intelligence: OpenAI's o3 Is Not AGI cites this paper.

Understanding and Benchmarking Artificial Intelligence: OpenAI's o3 Is Not AGI Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T20:46:08.961986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:46:08.961986Z digest=sha256:ad75ed0d1d60fe2af298c91b50c55b0122219dfb7e3350c9f4f20cbab3a9907f

Observation e398a153-d0d9-43ab-8532-405143c50ea3 · inbound

Psychometric-Based Evaluation for Theorem Proving with Large Language Models cites this paper.

Psychometric-Based Evaluation for Theorem Proving with Large Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T17:36:09.641336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:36:09.641336Z digest=sha256:cad1884503884f27643e77bd26b6083dbd331963069eb0e077196fd7e63f7720

Observation 19052245-579c-431b-b611-78b9e70a13fc · inbound

Probing for Arithmetic Errors in Language Models cites this paper.

Probing for Arithmetic Errors in Language Models Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:58:56.030900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:58:56.030900Z digest=sha256:20c34ff76291b23f1d4e749d9bf738e72a8867dccea09e4b2bf2011e20548b67

Observation e9279ed4-1816-4e1c-aade-e5498b578c27 · inbound

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs cites this paper.

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-17T04:34:02.116605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T04:32:48.056686Z digest=sha256:03fa63d43a3d5afa40f6a5372ca258a1147ce9b4a079026d27fbad2cd87241fa

Observation 85dcb22b-3584-46da-b5a3-b4bac1381a9d · inbound

Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms cites this paper.

Disentangling Mathematical Reasoning in LLMs: A Methodological Investigation of Internal Mechanisms Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:58:12.774484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T08:53:30.930960Z digest=sha256:1709fb16d16ca10e5aa0a5ffc99dbc593b8d1ba2600d325de59b39fc547ed5a9

Observation 3407cbb1-a55a-4171-8a89-326d92609d96 · inbound

From Execution to Education: A Bloom-Aligned Framework for Measuring Educational Control in LLMs cites this paper.

From Execution to Education: A Bloom-Aligned Framework for Measuring Educational Control in LLMs Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions

Reference 71

Resolution
verified exact
local_arxiv, observed 2026-07-10T13:57:06.845864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-10T13:49:17.343893Z digest=sha256:3b8158bf1c3b88f71e3817b1ae25f8bd3f5fe21f25935bbb79c5a38216e82b27