Pith. sign in

Paper Citation Record · LEDGER

ToolQA: A Dataset for LLM Question Answering with External Tools

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2306.13304.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.13304 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:45:45.925042Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

39
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 44c9d07d-b1b4-4e14-9e1f-d4cfcb900f9b · inbound

API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs cites this paper.

API-Bank: A Comprehensive Benchmark for Tool-Augmented LLMs ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:51:41.210477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T20:51:41.198767Z digest=sha256:e26f12beea87da01b651e453811124dcb91dd7a5afe5eda44182cfc6bca637d4

Observation 5851f73b-3105-4c26-94a8-e55a58e9eba2 · inbound

GAIA: a benchmark for General AI Assistants cites this paper.

GAIA: a benchmark for General AI Assistants ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-12T15:46:03.422284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-12T15:46:03.247029Z digest=sha256:7d0fd89aec5e07bf98468c0d30b67e5bb3a8a8dbd0853f93a4b2a91caf53cec6

Observation ae198f74-293a-4589-b2dc-19cc9699b27e · inbound

Large Language Models: A Survey cites this paper.

Large Language Models: A Survey ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 201

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:22:55.911464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T15:22:54.023279Z digest=sha256:5ef7a1c4540a2eff07fcd58ac0b16e966fb539d8664f49c5ba948f8ebc3bd642

Observation 3a23a044-8a08-4cd1-bd2f-9eb5911eb59b · inbound

Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey cites this paper.

Parameter-Efficient Fine-Tuning for Large Models: A Comprehensive Survey ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T11:32:36.961300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T11:32:36.738536Z digest=sha256:820b6fc7b9e6683e559b84df8c1e9227f4eb519b644ef132ebbc4a9528392e84

Observation bc9c5c3b-e457-41f0-8042-c3061a65302c · inbound

Learning to Ask: When LLM Agents Meet Unclear Instruction cites this paper.

Learning to Ask: When LLM Agents Meet Unclear Instruction ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-23T21:13:28.044365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-23T21:08:42.276002Z digest=sha256:ee9bdb9efe3aec11cdd99daa1486352892038318961f1f6218d8f59372c75df3

Observation e74afaba-005e-442f-af42-92dbc6dc0cc8 · inbound

Testing Uncertainty of Large Language Models for Physics Knowledge and Reasoning cites this paper.

Testing Uncertainty of Large Language Models for Physics Knowledge and Reasoning ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T18:25:29.443523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:25:29.443523Z digest=sha256:901ae2f77e2ea0f7fc08069aff2c3ce28dd2c1956ff318c3fe7eb8f47fd787fe

Observation 56fd37ba-2f00-48d8-b6e5-a51d06708945 · inbound

Reducing Tool Hallucination via Reliability Alignment cites this paper.

Reducing Tool Hallucination via Reliability Alignment ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T21:47:17.085295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:47:17.085295Z digest=sha256:51085aac4cbd92e63815288764d29523fde435bf58729edc00bd580f0f598333

Observation 79a0b9a5-63ee-4223-b57b-efb430a64e12 · inbound

PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines cites this paper.

PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-16T11:45:45.925042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:45:45.925042Z digest=sha256:eaaa4e6a4a9ecfeec6b510a2f8d0d1f78136523fd314cb33769705b8931a5ef5

Observation 4f430b31-00ab-4cef-ada7-9d96dea190c5 · inbound

A Framework for Testing and Adapting REST APIs as LLM Tools cites this paper.

A Framework for Testing and Adapting REST APIs as LLM Tools ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:26:54.627092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:26:54.627092Z digest=sha256:59ba2b3c4dc8f8afcd666954059ddbbfacfb15f61d0d0315ea05a515f5e2baa8

Observation adc348c2-4bb4-48c2-873d-02333adc2639 · inbound

Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges cites this paper.

Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T20:21:22.405524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:21:22.405524Z digest=sha256:c4f535c167abae15da0ec5ceac203a8cc8f83aadb59f480b281410ca3bb4f3fd

Observation 3e5d6cd4-2439-4fcc-9d6c-a6c20b2f61bf · inbound

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues cites this paper.

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:40.986835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:02:40.986835Z digest=sha256:3cca2d92f09b32e1a54c9eb7e40b8ce4f42ff845daeb4f3c911dfb67cbf70c63

Observation bc564f66-5f97-43eb-9382-55e696c5a97e · inbound

Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning cites this paper.

Beyond Accuracy: Unveiling Inefficiency Patterns in Tool-Integrated Reasoning ToolQA: A Dataset for LLM Question Answering with External Tools

Reference 7

Resolution
malformed identifier
arxiv_id, observed 2026-05-10T23:00:49.788025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T19:24:46.023589Z digest=sha256:4d593712991397bc7e024f5d5a9ba2f8a1ef228e4c3ee6ef0defdb8a79c0ae37