Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:10:04.649930Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 2 inbound Pith citation observations for arXiv:2509.04518.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:10:04.649930Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T21:54:13.090019Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T09:46:01.011609Z
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5a739158-c700-47d6-b97c-2a8a6e6ee523 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning xLAM: A Family of Large Action Models to Empower AI Agent Systems
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acdecbb0-7f67-498e-96ff-6bdf79cb8cb7 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Small Language Models: Survey, Measurements, and Insights
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8591ed5-cdcf-4658-9f78-2a8da7684f8a · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning A Comprehensive Survey of Small Language Models in the Era of Large Language Models: Techniques, Enhancements, Applications, Collaboration with LLMs, and Trustworthiness
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d357488-fbff-48f6-8760-0e1d122f737d · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning A Survey of Small Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 314604c8-6f0a-4502-825b-62a61616c468 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning A Survey on Large Language Model Based Autonomous Agents,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d0408b0-35b3-4889-bd7b-9c50d71362c5 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Report on a General Problem-Solving Program,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19602187-4f0b-43b8-8617-5db64f913ecd · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Recursive Functions of Symbolic Expressions and Their Computation by Machine, Part I,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 02c4f326-0cbf-45c8-938c-23847ffc2f9e · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Language Models are Unsupervised Multitask Learners,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a587a32d-68ca-43e2-bd58-4e71eeeb6eca · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning 'Alexa, Do You Know Anything?' The Impact of an Intelligent Assistant on Team Interactions and Creative Performance Under Time Scarcity
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed528a00-fba9-4f79-a27f-f7dfd58408f4 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning End-to- End Autonomous Driving: Challenges and Frontiers,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b35464ff-239f-42da-98bb-40e31cecb7a8 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning AI Agents That Matter
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b766be02-3661-4252-b449-f14034bf437b · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Toolformer: Language Models Can Teach Themselves to Use Tools,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ccdd332-9f88-4cd7-8a73-524320bf1074 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03d53b38-d8a9-4db5-a503-f4de6e79841b · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Training Language Models to Follow Instructions with Human Feedback,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4e28a8d-eeef-49f0-9ab0-e02b5a725ad6 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef53e933-8bd7-439b-86a5-6e4b5a864c74 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d523fb7a-35c5-448f-a7e1-2b1709459da8 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 831b30e6-4e91-4df4-aca9-717956e41689 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning TinyAgent: Function Calling at the Edge,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 535412ba-ab79-4f6c-bc3f-16386e6dc8a0 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Improving Small-Scale Large Language Models Function Calling for Reasoning Tasks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a8bba0-6122-44f4-ad60-afb7355bb878 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Qwen2.5 Technical Report
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 957abe72-1b9f-4cc3-9c9b-aee939aa7add · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning LLaMA: Open and Efficient Foundation Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7d9ffda-4c03-4d16-84f1-3a92e51592c7 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Qwen2.5 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f424498-7f7b-437d-862b-b2486113aa2b · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning Unresolved cited work
Reference 2017
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a58ba46-ed0d-4905-b6b0-141cc59415d8 · outbound
Advancing SLM Tool-Use Capability using Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9bff33a-e35d-4524-9579-123ad546bf1d · inbound
FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning Advancing SLM Tool-Use Capability using Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a4dd6d6-a61a-4d34-8ed9-cecbef6a54d1 · inbound
UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents Advancing SLM Tool-Use Capability using Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.