Pith. sign in

Paper Citation Record · LEDGER

Advancing SLM Tool-Use Capability using Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 2 inbound Pith citation observations for arXiv:2509.04518.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04518 v2

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:10:04.649930Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T21:54:13.090019Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T09:46:01.011609Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact2
  • verified fuzzy8
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5a739158-c700-47d6-b97c-2a8a6e6ee523 · outbound

This paper cites xLAM: A Family of Large Action Models to Empower AI Agent Systems.

Advancing SLM Tool-Use Capability using Reinforcement Learning xLAM: A Family of Large Action Models to Empower AI Agent Systems

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.975207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:00.978776Z digest=sha256:cec0abcb57f06d52fcaa5107041cd6b992530b8af564934f78e97e5a6f1be344

Observation acdecbb0-7f67-498e-96ff-6bdf79cb8cb7 · outbound

This paper cites Small Language Models: Survey, Measurements, and Insights.

Advancing SLM Tool-Use Capability using Reinforcement Learning Small Language Models: Survey, Measurements, and Insights

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:01.164458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:01.164458Z digest=sha256:d8747c0845eda62254c0100dd0d1a8f2036c95e33014df6bcfbc2b404dd7a83f

Observation d8591ed5-cdcf-4658-9f78-2a8da7684f8a · outbound

This paper cites A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness.

Advancing SLM Tool-Use Capability using Reinforcement Learning A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:01.491388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:01.491388Z digest=sha256:bb702386cbc9029617432c141930729c87b68270aa7ef37cfd5b3164ae60163e

Observation 0d357488-fbff-48f6-8760-0e1d122f737d · outbound

This paper cites A Survey of Small Language Models.

Advancing SLM Tool-Use Capability using Reinforcement Learning A Survey of Small Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:01.962412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:01.962412Z digest=sha256:17c9838e4ac711dcf803b790ecea2d0bad68f201eb8529d317f5d69128322bb5

Observation 314604c8-6f0a-4502-825b-62a61616c468 · outbound

This paper cites A Survey on Large Language Model Based Autonomous Agents,.

Advancing SLM Tool-Use Capability using Reinforcement Learning A Survey on Large Language Model Based Autonomous Agents,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:02.412340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:02.412340Z digest=sha256:82b128288a75cde925151ecdd1f7a3ca0b557a3faa4bfcf3c3660f4032dba71d

Observation 9d0408b0-35b3-4889-bd7b-9c50d71362c5 · outbound

This paper cites Report on a General Problem-Solving Program,.

Advancing SLM Tool-Use Capability using Reinforcement Learning Report on a General Problem-Solving Program,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.958735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:02.945525Z digest=sha256:6e934b6e43d2a3bfe414492932e5cb29fa909fe4b0d81bea7e5c720453cb38b2

Observation 19602187-4f0b-43b8-8617-5db64f913ecd · outbound

This paper cites Recursive Functions of Symbolic Expressions and Their Computation by Machine, Part I,.

Advancing SLM Tool-Use Capability using Reinforcement Learning Recursive Functions of Symbolic Expressions and Their Computation by Machine, Part I,

Reference 7

Resolution
verified exact
doi, observed 2026-08-05T11:10:04.939574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.110076Z digest=sha256:030995942958c245f17fd402013fd57f8ff4673bf00c6a287e4e5c6c6ed5e352

Observation 02c4f326-0cbf-45c8-938c-23847ffc2f9e · outbound

This paper cites Language Models are Unsupervised Multitask Learners,.

Advancing SLM Tool-Use Capability using Reinforcement Learning Language Models are Unsupervised Multitask Learners,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.942175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.208256Z digest=sha256:bd93f2fc7dd6788b0c308be8e6b45cabd2624ccd572ae0eb6b00f4d5916892f1

Observation a587a32d-68ca-43e2-bd58-4e71eeeb6eca · outbound

This paper cites 'Alexa, Do You Know Anything?' The Impact of an Intelligent Assistant on Team Interactions and Creative Performance Under Time Scarcity.

Advancing SLM Tool-Use Capability using Reinforcement Learning 'Alexa, Do You Know Anything?' The Impact of an Intelligent Assistant on Team Interactions and Creative Performance Under Time Scarcity

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:10:04.877547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.295382Z digest=sha256:830317dc4a7cdedf930ff164f44c92f4ecef8846134a6db5c771ad6452bcb2d6

Observation ed528a00-fba9-4f79-a27f-f7dfd58408f4 · outbound

This paper cites End-to- End Autonomous Driving: Challenges and Frontiers,.

Advancing SLM Tool-Use Capability using Reinforcement Learning End-to- End Autonomous Driving: Challenges and Frontiers,

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-05T11:10:03.382282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:03.382282Z digest=sha256:1df013bb14de55dade63d8eda8824d4d27fc11df912f10360d08e16e52a8827b

Observation b35464ff-239f-42da-98bb-40e31cecb7a8 · outbound

This paper cites AI Agents That Matter.

Advancing SLM Tool-Use Capability using Reinforcement Learning AI Agents That Matter

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:03.458069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:03.458069Z digest=sha256:572f766da8467b2799a1ac88675a6a357ff3741c746c9938600936e05c956e00

Observation b766be02-3661-4252-b449-f14034bf437b · outbound

This paper cites Toolformer: Language Models Can Teach Themselves to Use Tools,.

Advancing SLM Tool-Use Capability using Reinforcement Learning Toolformer: Language Models Can Teach Themselves to Use Tools,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.921779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.567181Z digest=sha256:52c9fb8586a7731d25bcfb154440bb183637092daa60d8384f910e16dbe4a13a

Observation 4ccdd332-9f88-4cd7-8a73-524320bf1074 · outbound

This paper cites ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs,.

Advancing SLM Tool-Use Capability using Reinforcement Learning ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.801284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.642778Z digest=sha256:3a1e5fc9ccbb37bab71c6ae5745175b844bf652a2efef1c62cb1647888abdf43

Observation 03d53b38-d8a9-4db5-a503-f4de6e79841b · outbound

This paper cites Training Language Models to Follow Instructions with Human Feedback,.

Advancing SLM Tool-Use Capability using Reinforcement Learning Training Language Models to Follow Instructions with Human Feedback,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.645364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.751242Z digest=sha256:34e08524f3a2d26e295e05d1856ea25f69a27ec60b173c77eac416e549ab9741

Observation e4e28a8d-eeef-49f0-9ab0-e02b5a725ad6 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models,.

Advancing SLM Tool-Use Capability using Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.517883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:03.843310Z digest=sha256:7a992d2bc56ee7e20c2ec8551e4a0f4518dc6cf3948fcabacddfd101e3291141

Observation ef53e933-8bd7-439b-86a5-6e4b5a864c74 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Advancing SLM Tool-Use Capability using Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:04.079194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:04.079194Z digest=sha256:61f514afbb0603645e02b388d11210043c37471f3640042b7ee6a2d342254018

Observation d523fb7a-35c5-448f-a7e1-2b1709459da8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Advancing SLM Tool-Use Capability using Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:04.188526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:04.188526Z digest=sha256:9cbf727663b8266029c105a9264440c177464ea02d8f27618897cbe492537585

Observation 831b30e6-4e91-4df4-aca9-717956e41689 · outbound

This paper cites TinyAgent: Function Calling at the Edge,.

Advancing SLM Tool-Use Capability using Reinforcement Learning TinyAgent: Function Calling at the Edge,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T11:10:05.333086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:04.304885Z digest=sha256:34fabc5961967bc90a2af55fd8533def927ad465d184d92276371ccdf22fffda

Observation 535412ba-ab79-4f6c-bc3f-16386e6dc8a0 · outbound

This paper cites Improving Small-Scale Large Language Models Function Calling for Reasoning Tasks.

Advancing SLM Tool-Use Capability using Reinforcement Learning Improving Small-Scale Large Language Models Function Calling for Reasoning Tasks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:04.397873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:04.397873Z digest=sha256:3dbc6a2c636adfcbc242e1078466ccde81b44f8f99a13589aeac6f19cdded47e

Observation 66a8bba0-6122-44f4-ad60-afb7355bb878 · outbound

This paper cites Qwen2.5 Technical Report.

Advancing SLM Tool-Use Capability using Reinforcement Learning Qwen2.5 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:04.469320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:04.469320Z digest=sha256:7bbaecd17825226fe6e97e153da7236f329bf1f33fc9a699f1a51bd8c9b43b33

Observation 957abe72-1b9f-4cc3-9c9b-aee939aa7add · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Advancing SLM Tool-Use Capability using Reinforcement Learning LLaMA: Open and Efficient Foundation Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:04.649930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:04.649930Z digest=sha256:49402d62c713dc89cdc2baf72f0f366f7a48e917c41209f5c11a2d2cb73b86a3

Observation f7d9ffda-4c03-4d16-84f1-3a92e51592c7 · outbound

This paper cites Qwen2.5 Technical Report.

Advancing SLM Tool-Use Capability using Reinforcement Learning Qwen2.5 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:04.581497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:04.581497Z digest=sha256:c7acca67db640d71b4187c3fddd0dde4dc52cc9b4650f00dce6cf9b8e470716a

Observation 6f424498-7f7b-437d-862b-b2486113aa2b · outbound

This paper cites an unresolved cited work.

Advancing SLM Tool-Use Capability using Reinforcement Learning Unresolved cited work

Reference 2017

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:10:05.411430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T11:10:04.202015Z digest=sha256:9e4d5c0fce66bfb4ca340b75add73a28f28609e1c1621deefec931690975fa17

Observation 5a58ba46-ed0d-4905-b6b0-141cc59415d8 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Advancing SLM Tool-Use Capability using Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T11:10:03.953925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:10:03.953925Z digest=sha256:077331e0171a6ebb5c6edceb8c21ac66f04b92d7d02e6d0161f30ac2acad733e

Pith citing papers

Observation d9bff33a-e35d-4524-9579-123ad546bf1d · inbound

FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning cites this paper.

FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning Advancing SLM Tool-Use Capability using Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T21:54:13.090019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:54:13.090019Z digest=sha256:2c9f771fc5a998105c860eb876154866cc2daf695dee7e516475243c3cc9199b

Observation 2a4dd6d6-a61a-4d34-8ed9-cecbef6a54d1 · inbound

UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents cites this paper.

UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents Advancing SLM Tool-Use Capability using Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:46:01.014476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:51:22.508030Z digest=sha256:2144bbd162543bea506e36cd5843e2127563a07f89964472d96def2bd8b8dc7c