Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:21:22.416314Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 2 inbound Pith citation observations for arXiv:2505.13328.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:21:22.416314Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T17:56:47.089599Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T17:56:51.738856Z
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bfd89ea3-3418-41db-906e-68600786822b · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 976a5420-ddfe-4684-88cc-c55c45e6e571 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Qwen Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f001b702-622b-4863-bf3a-035bf3b52f45 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0434f184-a0fe-44db-a97c-f23f13e81ed4 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97cf81e5-b89d-4e27-ac98-cae7a2a73a76 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Wizard of Wikipedia: Knowledge-Powered Conversational agents
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c122317-852f-41a7-af9d-acd353c862b5 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8a983d1-90a0-4947-9e68-45745d6dbc24 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges LLM as OS, Agents as Apps: Envisioning AIOS, Agents and the AIOS-Agent Ecosystem
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 233a01b5-6ff7-40ec-b1ba-b3cf6dad6982 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8152d58d-6f9f-46a9-addc-ef5a90babe9e · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 93261e84-8f4f-4cb4-bee5-3f8229532169 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Planning, Creation, Usage: Benchmarking LLMs for Comprehensive Tool Utilization in Real-World Complex Scenarios
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45bb2c7e-e636-4759-b1d1-b7414dccb115 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebd0970c-d357-4967-a142-175dc3457a0c · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d27af6e-1617-4e5d-b691-ed7dd0ab89b6 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Mistral 7B
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a89fcbfc-176f-4ec6-81de-d3c14286f919 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c69dc4e2-cbb7-42e9-8617-da2224961602 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1c4b58d0-74b6-49d1-aa6f-29a04d5a8ecd · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3363668e-9a48-4983-87b3-8f058e75f83e · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c291d8c-cbf8-440a-bd2c-07e2fa8c7a0f · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Code as Policies: Language Model Programs for Embodied Control
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 470239f6-ba28-40e9-bb14-c5994ac86121 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges AgentBench: Evaluating LLMs as Agents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1873e3ea-ba02-4f5c-a9c5-b8504ac0da62 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges ToolSandbox: A Stateful, Conversational, Interactive Evaluation Benchmark for LLM Tool Use Capabilities
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8329736-4908-4354-9036-50ef0c05b8ce · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Chameleon: Plug-and-Play Compositional Reasoning with Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e12d8fcb-21d4-4a51-9490-6da9ccbc8024 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges AgentBoard: An Analytical Evaluation Board of Multi-turn LLM Agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89341ce8-e9a6-4a0e-8d97-5bea7134ae7b · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges GAIA: a benchmark for General AI Assistants
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2ca764e-9509-47ba-ad41-7913336dff25 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges WebGPT: Browser-assisted question-answering with human feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cff8c72-4ca8-40fe-988c-902fb042ac3d · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Gorilla: Large Language Model Connected with Massive APIs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e576ddf2-50af-4364-ac8a-b7622f3a0766 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Few-shot Natural Language Generation for Task-Oriented Dialog
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7217ea47-1adf-4e73-987c-fb16d574af10 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d0129e7a-bc08-47d9-99b9-e7527832d429 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 028a17bb-279e-49fd-90f0-4d6585a659c1 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Tool Learning with Foundation Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a951ee21-d661-4003-855b-0dcc2f389231 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 390fa33f-00f7-476e-b2d5-a0d400046213 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 71c054eb-7db8-4503-af45-66ddefdeec29 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2e76e55-c5e4-47e2-b422-8af1050f183b · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Dialog2API: Task-Oriented Dialogue with API Description and Example Programs
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7d1599-36f0-43bb-aea7-e5fe910c9b2a · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Cognitive Architectures for Language Agents
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6d19186-4815-4322-b422-5254acc57821 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b5e9d7ac-5cb1-482e-a8ce-676a128a783e · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7a12258-2f5f-4179-85ae-7ecc4f5a7048 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges CharacterEval: A Chinese Benchmark for Role-Playing Conversational Agent Evaluation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deddd860-eb66-406f-8e85-845a06c863db · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ab012a2b-fcda-4aec-b796-48417b1d9ad4 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 99db0743-4b5f-4621-b857-4e70b2177cfc · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Pan, and Kam-Fai Wong
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ceeb9ad-05f4-47ca-8a5b-c565493e841e · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0ecedaed-9c0a-450f-9912-e53c351e171b · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges A Survey of the Evolution of Language Model-Based Dialogue Systems: Data, Task and Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e29f278-418c-49e0-ad6f-1b49d6e7804d · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81ecf18d-2150-41a5-be36-529c63b84e47 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Pan, and Kam-Fai Wong
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 759242c7-8a97-4167-885d-c296addf5fa3 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c9f7a41-f11c-4b04-9037-dea4e97f8840 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2099b230-e2c8-494b-8a8f-8fc1a85b6b63 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4916298-04a0-4710-b894-8bac46be5fef · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges RoleLLM: Benchmarking, Eliciting, and Enhancing Role-Playing Abilities of Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7c7d5a7-d410-46cd-b7f3-337db7b24c12 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8f4ca26-2ce3-461f-8aaf-fc19f807f188 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e6b3a31-4483-40ae-9fa7-d7be07bfd267 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5faf31fe-b4b0-4e01-a865-5d3f436bd377 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cab0b35d-329d-45fa-9469-cd064975dac4 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ea6efc9-9cc3-4485-bab7-e3c71d23b362 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges CharacterGLM: Customizing Chinese Conversational AI Characters with Large Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1490c7d-4df1-40b7-b01f-4a1550928802 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adc348c2-4bb4-48c2-873d-02333adc2639 · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges ToolQA: A Dataset for LLM Question Answering with External Tools
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89db7830-05bd-45a3-bcf7-d1af6c9e01eb · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges online" 'onlinestring :=
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0776f5d2-6d41-49b7-8761-ea38604b646b · outbound
Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges write newline
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24699e33-12a6-4b9e-89d9-cf52684af9ce · inbound
Experience-Evolving Multi-Turn Tool-Use Agent with Hybrid Episodic-Procedural Memory Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b16a860-d196-4d13-ae6c-60fa27126767 · inbound
WeClawArena: An Auditable Sandbox and Benchmark for Cross-User Agents Collaboration and Security in Human-Centered Agent Networks Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.