Pith. sign in

Paper Citation Record · LEDGER

On the Tool Manipulation Capability of Open-source Large Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2305.16504.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.16504 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:00:19.615486Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T17:47:17.727414Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 141acb86-5080-4989-8981-cd27ad4641eb · inbound

ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases cites this paper.

ToolAlpaca: Generalized Tool Learning for Language Models with 3000 Simulated Cases On the Tool Manipulation Capability of Open-source Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:03:48.486935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T23:03:48.426204Z digest=sha256:4b5f612eaaa95b128d4cf399b5b66b3d5830b29f321be1f946d5891d6829ef48

Observation f54a43fb-c740-4ab5-a05f-6f83d4a42093 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models On the Tool Manipulation Capability of Open-source Large Language Models

Reference 222

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:38.997640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:d03780f2ea23096a0bd5a8bd006c0b72933b3c5739692140174317b9cf03dfda

Observation 223db9e6-9553-44c1-83e0-08db59ff02f7 · inbound

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems cites this paper.

From Standalone LLMs to Integrated Intelligence: A Survey of Compound Al Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 203

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:52:16.379714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T11:49:36.574471Z digest=sha256:011b59a72d8d54f39d5e5641bc991839dc5b23cab5767a5304644b9860db6582

Observation e27cad62-82a1-4cf2-8979-786243445a7f · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey On the Tool Manipulation Capability of Open-source Large Language Models

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.615486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.615486Z digest=sha256:a182e72f1bd09ca90d4dce2409cff7ea84cbd796913497a59fe15cd889583918

Observation 9a216b23-d2b7-4c7b-a3e1-e26f67b2ee8d · inbound

CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios cites this paper.

CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios On the Tool Manipulation Capability of Open-source Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:44:04.866566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:44:04.866566Z digest=sha256:30dcff6877ef6796d929ed2f87b429f3480cc50e79fe94bb8e5d89cfba4a277c

Observation 5321c14a-8d40-4db5-a641-3352980a350c · inbound

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues cites this paper.

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues On the Tool Manipulation Capability of Open-source Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:40.749933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:02:40.749933Z digest=sha256:70ec1ab93de791a2b5213fc1445c4d8611d25178f1efd2e67ace97594d46c1d6

Observation 06320fdd-fbd6-44de-969d-065a372ead7a · inbound

MassTool: A Multi-Task Search-Based Tool Retrieval Framework for Large Language Models cites this paper.

MassTool: A Multi-Task Search-Based Tool Retrieval Framework for Large Language Models On the Tool Manipulation Capability of Open-source Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:43.855623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:22:43.855623Z digest=sha256:4d812bba2e1c2b1363a2e68f184b735a301987e4ccc165555edfb9bad92e9c3d

Observation a3b943cd-2ceb-45d2-bf5a-b805297b5a7e · inbound

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis cites this paper.

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis On the Tool Manipulation Capability of Open-source Large Language Models

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:40:51.380018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T00:37:11.945418Z digest=sha256:bde840470bda41c4a41bdd3865ffa1b651b18584e3cfecfdcb01585176519ea0

Observation 56fe74af-f73f-4212-a055-c89c49c0ca8a · inbound

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems cites this paper.

A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 106

Resolution
verified exact
arxiv_id, observed 2026-05-15T23:21:42.199107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T23:21:42.029285Z digest=sha256:921480d7a35b418d770be4761923efeffd7d9d0f38f323e87346c03ea9c7e079

Observation a8db8f16-c010-416d-b034-7cb28b28d022 · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 260

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:18:20.242995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:7e0038daffe2aa707d067a885276008b7f0b2f1a5dba8876e72b9702b02cfbd3

Observation dd321c8a-80c4-48f9-8e85-7e8113c85983 · inbound

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents cites this paper.

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T23:41:45.208041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:41:45.208041Z digest=sha256:900604ade693cf5e2caf82124a99a91f83512e4ec6757065b35e6942cd3f7427

Observation c8d96b6f-6b18-422a-afd4-35ab12789168 · inbound

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents cites this paper.

Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:35:51.573419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:59:59.383781Z digest=sha256:6f5fde0ef81ad71c4fe37ea8a38ee8b59234847910fa348497dcfcd7ac0ea11d

Observation 56c9e580-ad0c-468d-bb26-9bbba19359ab · inbound

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation cites this paper.

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation On the Tool Manipulation Capability of Open-source Large Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:46:24.179373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T14:01:23.894966Z digest=sha256:1435df225c1ef956ae5b220f07965171a4dec1925a81fd91fea60e6dd9f77926

Observation cd26da40-5c61-4cbd-8219-8a6f79338f7f · inbound

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions cites this paper.

Intent2Tx: Benchmarking LLMs for Translating Natural Language Intents into Ethereum Transactions On the Tool Manipulation Capability of Open-source Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:31:29.337321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T05:49:57.957009Z digest=sha256:d265d55b827c8acbd16f2086fe76b9f9b0985a5bb143b2e3ddc45a789dad9686

Observation c9711691-ec57-4762-b5f1-c871a6dac7de · inbound

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems cites this paper.

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:21:27.819884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:19:32.031454Z digest=sha256:edaf4896cfcb1baa9d465f2cff87e269daf1c86f81d843e5adce0c6b66451747

Observation 029be727-401f-4c00-aa03-e8c5953c8e2a · inbound

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents cites this paper.

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents On the Tool Manipulation Capability of Open-source Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.255056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T22:27:16.974169Z digest=sha256:9efefc99a11d51f1c45a7973d0b46453b9dddaeac0e424d287d69f5dc1f1be40

Observation 840da58a-95ab-4cc5-8812-1948d7c2153e · inbound

Continual Model Routing in Evolving Model Hubs cites this paper.

Continual Model Routing in Evolving Model Hubs On the Tool Manipulation Capability of Open-source Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:03:24.148298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T11:57:15.894309Z digest=sha256:ebf796cb3924716ed6ec180d1f09832e7cd85db405405d50514dc8a017e633d8

Observation 9fb5f490-80ad-41e8-90a9-6f2d5f558164 · inbound

Capability Self-Assessment: Teaching LLMs to Know Their Limits cites this paper.

Capability Self-Assessment: Teaching LLMs to Know Their Limits On the Tool Manipulation Capability of Open-source Large Language Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.682649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:25:50.579196Z digest=sha256:c79a9d63ea87bc4e5a841df6292ec8694652a479037aab3e02aeafde6d77cd0f

Observation 32594cdc-c444-4477-96cc-3f6f2062ddec · inbound

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems cites this paper.

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems On the Tool Manipulation Capability of Open-source Large Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:47:17.728995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T21:53:37.616447Z digest=sha256:dfa0b8bc0f866bd49b49f568fe8065014006d5e7e505701aee3c20b684ae614e

Observation ba7935e2-746e-4d98-ab8a-713c6f842280 · inbound

A Formal Hierarchical Architecture for Agentic Orchestration with Stack-Based Execution and Lazy Discovery cites this paper.

A Formal Hierarchical Architecture for Agentic Orchestration with Stack-Based Execution and Lazy Discovery On the Tool Manipulation Capability of Open-source Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T06:48:32.645976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:48:32.645976Z digest=sha256:6945d54085a9b79a686a1a75c54864055a567aefe224fa72f016b95b194c6885