Pith. sign in

Paper Citation Record · LEDGER

TheoremQA: A Theorem-driven Question Answering dataset

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2305.12524.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.12524 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:22:24.562397Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3dcfd77e-f6fd-484c-820f-60e49dc0199c · inbound

Detecting Language Model Attacks with Perplexity cites this paper.

Detecting Language Model Attacks with Perplexity TheoremQA: A Theorem-driven Question Answering dataset

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T14:02:26.851987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T14:02:26.784965Z digest=sha256:30f081fca685e4548c3d55cf7050779e01c2b0100207e885cf4e9f425bab56e1

Observation c56ec44e-dcdf-4312-9fd8-4f0e53bc1777 · inbound

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning cites this paper.

MAmmoTH: Building Math Generalist Models through Hybrid Instruction Tuning TheoremQA: A Theorem-driven Question Answering dataset

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T23:46:39.578030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-17T23:46:39.330438Z digest=sha256:d6c84d7dac706970c090cb347f7ee776f776e1180caf2fb23257bc3a8eccb746

Observation 5a431d61-daf8-4a7e-897a-e5c964430a73 · inbound

PhySense: Principle-Based Physics Reasoning Benchmarking for Large Language Models cites this paper.

PhySense: Principle-Based Physics Reasoning Benchmarking for Large Language Models TheoremQA: A Theorem-driven Question Answering dataset

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:22:24.562397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:22:24.562397Z digest=sha256:45b28e25023c47bf3054cf9283b53c800261f5b2b698ad3ff4eba7131efa555f

Observation db6db4af-9926-4f46-9e83-3b382a3516cf · inbound

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos cites this paper.

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos TheoremQA: A Theorem-driven Question Answering dataset

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:22:55.808811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:22:55.808811Z digest=sha256:1da557a54b3e78767ce678464bd330d0739a15c7abfd18545df0d73c50bf185d

Observation 544c9ab7-187e-4720-86cb-1ddfdb39c3ef · inbound

AI4Research: A Survey of Artificial Intelligence for Scientific Research cites this paper.

AI4Research: A Survey of Artificial Intelligence for Scientific Research TheoremQA: A Theorem-driven Question Answering dataset

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:12.276433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:12.276433Z digest=sha256:fff5f820089eb61889dfceb0266c7186a57da4b99cf67c940b2c7f1306686350

Observation 73dc1681-ec6e-4b8f-9a9b-84e118096ee2 · inbound

A Simple "Try Again" Can Elicit Multi-Turn LLM Reasoning cites this paper.

A Simple "Try Again" Can Elicit Multi-Turn LLM Reasoning TheoremQA: A Theorem-driven Question Answering dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:14:06.336764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:14:06.336764Z digest=sha256:18916fb6c8830223f1df48c55024259f434ad2dd7c31b6da4154c4812f0d5697

Observation badd2157-376d-4368-8576-60a08c70ce9c · inbound

Technical Report of TeleChat2, TeleChat2.5 and T1 cites this paper.

Technical Report of TeleChat2, TeleChat2.5 and T1 TheoremQA: A Theorem-driven Question Answering dataset

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:22.133785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:43:22.133785Z digest=sha256:db916c894162e20dd61638d410acaa93fa9bc0707d0976d7d9955e60a9702660

Observation 07e29574-872c-411d-8ed4-ad2c22df3d9d · inbound

CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning cites this paper.

CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning TheoremQA: A Theorem-driven Question Answering dataset

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T13:53:32.793396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:53:32.793396Z digest=sha256:02447efbb71c927557453536fc01e2eb089633ab1fe5181f4d2a78b9add6cccd

Observation b780332e-3272-4f9f-9140-df45cd86e9c8 · inbound

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics cites this paper.

StatEval: A Comprehensive Benchmark for Large Language Models in Statistics TheoremQA: A Theorem-driven Question Answering dataset

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T10:36:23.048036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T10:36:23.048036Z digest=sha256:cc8c2f0bc3c756e0ab6cf4df6a30b03ec9a0ed09f5e20643d4caf41c276808ee

Observation c0545efa-1a4e-4f6a-8a2b-a39b66788a1a · inbound

Coupled Variational Reinforcement Learning for Language Model General Reasoning cites this paper.

Coupled Variational Reinforcement Learning for Language Model General Reasoning TheoremQA: A Theorem-driven Question Answering dataset

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T16:43:54.710371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:43:54.710371Z digest=sha256:e521ad66e99e0eccd4bb486242493f6f77b22847744ffeb3ec74cbc07159f335

Observation 496e4370-6c12-4101-855b-9466cceec120 · inbound

Empirical Characterization of Inference-Time Elicited Probability Transformations in Large Language Models cites this paper.

Empirical Characterization of Inference-Time Elicited Probability Transformations in Large Language Models TheoremQA: A Theorem-driven Question Answering dataset

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-02T20:21:54.834919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:21:54.834919Z digest=sha256:e384571c065a6d3fc96ec5a884613f4c7659668c2174a516e0ccfa016f69f7d5

Observation e4124a06-c1d1-4248-b5d1-0b04d08da3ca · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning TheoremQA: A Theorem-driven Question Answering dataset

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:52:59.463228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T16:51:48.705876Z digest=sha256:f5becb2dd8e19232cbc5db7f068444672af28227937f88bf78172d7c4f450112

Observation 0ac21af9-42bb-4318-b6e2-3946858e5435 · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning TheoremQA: A Theorem-driven Question Answering dataset

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T12:10:53.720348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:10:53.720348Z digest=sha256:9e36e9b75de2022c53394fffca79ad6fc120b122257993044bbe2117b1f6f305

Observation e4e1e23c-db95-414c-a114-ae0def66b3db · inbound

Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs cites this paper.

Transforming External Knowledge into Triplets for Enhanced Retrieval in RAG of LLMs TheoremQA: A Theorem-driven Question Answering dataset

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:05.518220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:42:59.063869Z digest=sha256:2a62c6e0994582367fcb999115d95e3bd292322deb727f5dd45499d183f431d1

Observation 0d0eb227-86e3-4549-a74b-8c292997a0ab · inbound

GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning cites this paper.

GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning TheoremQA: A Theorem-driven Question Answering dataset

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:14:46.361796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T00:14:29.531510Z digest=sha256:00dbcdac7f1a74171c9abf56b32723f87bc41b164142744f4efcbd67befa3752

Observation 9023fc29-e263-4504-b6c1-533bad79485e · inbound

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces cites this paper.

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces TheoremQA: A Theorem-driven Question Answering dataset

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:01:18.568730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T02:57:15.521594Z digest=sha256:22d2e8fe75356dc0a2045682d426ca5af6741136ba55ad9c7c93852e0bea2f26

Observation c8ce4cd7-db9d-439b-940a-9042190d010f · inbound

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics cites this paper.

Re$^2$Math: Benchmarking Theorem Retrieval in Research-Level Mathematics TheoremQA: A Theorem-driven Question Answering dataset

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:16.053978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:07:39.648269Z digest=sha256:87b9e7098f4e08035c15fa35948331585921dfdab8f4dfe8a7da62c2f49be618

Observation 92230905-ca36-4c93-b155-946231accca7 · inbound

FIND: Toward Multimodal Financial Reasoning and Question Answering for Indic Languages cites this paper.

FIND: Toward Multimodal Financial Reasoning and Question Answering for Indic Languages TheoremQA: A Theorem-driven Question Answering dataset

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:42:58.871798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:24:46.503457Z digest=sha256:b79692c338f5dd1806bb75e15663ea5b195274cc2bf16f00aab5a3c28e1ec983

Observation 4abd39e9-1b89-4c49-95e0-aafdeeff64b6 · inbound

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization cites this paper.

DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization TheoremQA: A Theorem-driven Question Answering dataset

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:02:50.300057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T23:16:49.358792Z digest=sha256:bf58ec90183ee89c7959cb46282befcd21e95230cc2ca11a8c79954bfe89941c

Observation e27ac166-aa33-4c50-a297-6849f3dd9cca · inbound

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents cites this paper.

EEVEE: Towards Test-time Prompt Learning in the Real World for Self-Improving Agents TheoremQA: A Theorem-driven Question Answering dataset

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:47:38.824816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T13:34:32.750618Z digest=sha256:51a658d32d03caa6ff91889d59c571e8b29e284413956bb9af08cdc48255a1e8

Observation b03d169c-a3c8-4fdc-b97e-0aa2679f2167 · inbound

Structured Thoughts For Improved Reasoning And Context Pruning cites this paper.

Structured Thoughts For Improved Reasoning And Context Pruning TheoremQA: A Theorem-driven Question Answering dataset

Reference 59

Resolution
unresolved
no resolver link, observed 2026-07-14T12:08:05.502310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T12:08:05.502310Z digest=sha256:e2aa5fc3066c553fcfae7ffc8d828b5ffbc327d027e84576fb43e619f67db031