Pith. sign in

Paper Citation Record · LEDGER

Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.01763.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.01763 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:00:46.353616Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T19:17:18.694934Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6b1fa5e0-0118-494e-9098-ed4ab3e378e0 · inbound

LongFuncEval: Measuring the effectiveness of long context models for function calling cites this paper.

LongFuncEval: Measuring the effectiveness of long context models for function calling Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T05:00:46.353616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:00:46.353616Z digest=sha256:ef921e454ea3198c9a3e8fa895bc24eb34cfa5cd0c79e05bc066aa56ad29b3f6

Observation 7014bd2e-460c-4e7f-a7b6-cef93a233258 · inbound

Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows cites this paper.

Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:06:52.857377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T07:05:43.392997Z digest=sha256:584993f763f8839ed46b288f1e49137b35ac92d55604a0056f28998568e30235

Observation 36e07b27-076f-4196-8584-e80f908b9f5b · inbound

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval cites this paper.

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:05:33.916123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-08T19:04:53.639217Z digest=sha256:071493a346e236aaeaf6083802fd9e72991106fcab970e32f06abbee99b39ebe

Observation 7dfb1caa-3a67-46cb-a6ee-f03eb3ba4714 · inbound

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval cites this paper.

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-01T00:45:12.445369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-01T00:25:46.507401Z digest=sha256:7012a900cadb4c7a32f6aa384bd69a8603b2d3381852e19d34db8017e12e8de9

Observation 4ea31ecc-7d2f-484e-84f1-f6ce1ac46986 · inbound

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval cites this paper.

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T05:24:47.420184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T05:24:47.420184Z digest=sha256:987a22370127e369a0d4c2df3aa2380708c94e8198a580486f39797310e8acfe

Observation 06f3276b-f920-4f3e-9f87-fc2640f8762c · inbound

SkillRet: A Large-Scale Benchmark for Skill Retrieval in LLM Agents cites this paper.

SkillRet: A Large-Scale Benchmark for Skill Retrieval in LLM Agents Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:36:07.845804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T11:39:06.544414Z digest=sha256:6ad208eba262c9880695ca024a9e2617f4f512e4de0add291f2128299860e3e2

Observation 735b2ee2-08f2-47f5-8c4a-c2f88c2a9eb6 · inbound

SkillDAG: Self-Evolving Typed Skill Graphs for LLM Skill Selection at Scale cites this paper.

SkillDAG: Self-Evolving Typed Skill Graphs for LLM Skill Selection at Scale Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:46:29.320340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T10:35:35.230305Z digest=sha256:a5f276dfef106ea67fec1fb7d54387561ddfd3033929cf30e075a36df960d7af

Observation 05684f3d-2b46-4a15-83d1-de57d6894af7 · inbound

Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents cites this paper.

Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:17:18.696398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T21:32:06.237095Z digest=sha256:da02d9aa2a64094f296a7b7be246e27f8d9b53199b179a276ce098f09767486a

Observation eee7bfb1-ae77-46c7-87bb-629c3b8e1596 · inbound

Field Aware Agent Skill Retrieval cites this paper.

Field Aware Agent Skill Retrieval Retrieval Models Aren't Tool-Savvy: Benchmarking Tool Retrieval for Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:15.211168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:15.211168Z digest=sha256:95c83a0385d5114cd857f2491e88fe6d980d520f92dcab23d73f5a8459aff3e8