Pith. sign in

Paper Citation Record · LEDGER

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs

As of 16 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2608.10042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10042 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:18:07.485591Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact8
  • verified fuzzy13
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a0f803dc-53ec-4aa4-86b0-bd2376cd6061 · outbound

This paper cites Tuesday Retail Notes.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Tuesday Retail Notes

Reference 1

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:18:07.934741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.440337Z digest=sha256:3aa32337f43f2d70d3064243fd3bc079a64f4a073dcdd5c37452735868588e86

Observation db2e51a4-4383-45c8-bf7e-ea8884e0005d · outbound

This paper cites Personalized Language Modeling from Personalized Human Feedback.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Personalized Language Modeling from Personalized Human Feedback

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.672677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.672677Z digest=sha256:dc39b9be0644df01762b95c525d62434147a15908d94668863c04d287d5830e8

Observation 517e448d-53ec-4278-a7ac-fb773b0f178b · outbound

This paper cites Aligning LLMs by Predicting Preferences from User Writing Samples.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Aligning LLMs by Predicting Preferences from User Writing Samples

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:18:08.623168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:06.724740Z digest=sha256:bad292f040f4c813e15f43fd28223df589cf18744af778c6f62aa4791b8f8cbb

Observation a7b0ef19-e581-4183-987f-e346d5b353c1 · outbound

This paper cites API-bank: A comprehensive benchmark for tool-augmented LLMs.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs API-bank: A comprehensive benchmark for tool-augmented LLMs

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.874925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:06.774765Z digest=sha256:2a5ae5ec8a71d29917ea2f75c8753d599070f56cee56ab9e29e1d1b970faed7e

Observation ad1d714d-76d2-469d-b3a6-25bc860e3ea8 · outbound

This paper cites Benchmarking LLM Tool-Use in the Wild.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Benchmarking LLM Tool-Use in the Wild

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.864763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.864763Z digest=sha256:3b536ca92eb598cbcdaf38c5569557c3e875e1a440127dd5796835f343f67350

Observation 1dd23d7c-8bac-4165-8e68-e4205f2b0193 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.914755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.914755Z digest=sha256:757c334c06c4cc76cfb9ff888c2469bffbf2971fd356e12d79520e64fa0f840f

Observation c6b40833-3471-441e-8ff9-c6fa643a70ea · outbound

This paper cites Advancing and Benchmarking Personalized Tool Invocation for LLMs.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Advancing and Benchmarking Personalized Tool Invocation for LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.995679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.995679Z digest=sha256:c75fc94339b0d1f21a41afc351cb4f8b0b710901a75ae312479ec0f8908015aa

Observation a72393e1-6011-4523-b149-226db075730a · outbound

This paper cites Tool- spectrum: Towards personalized tool utilization for large language models.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Tool- spectrum: Towards personalized tool utilization for large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.655761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.054748Z digest=sha256:6d41ba6a80cfc6ace1b7700a32f38f380c181555b5a52ab47f5f5f22d05b7ccd

Observation c3dbc6e2-2c05-4cd2-b0fd-6ffb7a7af9b1 · outbound

This paper cites doi:10.18653/v1/2026.acl-long.370.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs doi:10.18653/v1/2026.acl-long.370

Reference 12

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.767881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.084830Z digest=sha256:8f12361ff0d2f891d0927e1cc67b85a7aa08fd605139e510d9a170f30076f187

Observation 7ee59956-cf73-4f46-b6a2-0b4c852d1d37 · outbound

This paper cites Fingertip 20k: A benchmark for proactive and personalized mobile llm agents.arXiv preprint arXiv:2507.21071,.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Fingertip 20k: A benchmark for proactive and personalized mobile llm agents.arXiv preprint arXiv:2507.21071,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.135485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.135485Z digest=sha256:e736c9f3d3d725c6e7bfb16795ee83c787bdbdeca8bba7459a940857606796d0

Observation 979fbe54-b07c-42de-b2c6-13e77d830895 · outbound

This paper cites Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.145058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.145058Z digest=sha256:596484068ffd68c48899f292c45374c979784d7010aa9e3bba8ba80b45aa4341

Observation 3d496b7a-8c64-40ac-b6e0-1805b5055a90 · outbound

This paper cites Me-agent: A personalized mobile agent with two-level user habit learning for enhanced interaction.arXiv preprint arXiv:2601.20162, 2026a.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Me-agent: A personalized mobile agent with two-level user habit learning for enhanced interaction.arXiv preprint arXiv:2601.20162, 2026a

Reference 15

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:18:08.401262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.150631Z digest=sha256:d57db262a81ef6f455d0cd5486dba223ebf1341ea331439ad4db3ce59738d548

Observation f8c2895f-6e01-46f0-a2c0-07fe6acf8e51 · outbound

This paper cites ValuePilot: A Two-Phase Framework for Value-Driven Decision-Making.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs ValuePilot: A Two-Phase Framework for Value-Driven Decision-Making

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:18:08.286973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.165600Z digest=sha256:e3860a066b5eaea114d4acc00983221b8ea1c0348f927789cec51cc5842322bc

Observation 68011b94-2c97-44dd-8284-0a887c959a75 · outbound

This paper cites Shopsimulator: Evaluating and exploring rl-driven llm agent for shopping assistants.arXiv preprint arXiv:2601.18225, 2026b.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Shopsimulator: Evaluating and exploring rl-driven llm agent for shopping assistants.arXiv preprint arXiv:2601.18225, 2026b

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.173493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.173493Z digest=sha256:bac329106df34cce17083be49d546b97e3411ca55383211c1d47c908527464ec

Observation 6932b77f-8bd6-4fed-9041-7eb330819605 · outbound

This paper cites PersonaLLM: Investigating the abil- ity of large language models to express personality traits.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs PersonaLLM: Investigating the abil- ity of large language models to express personality traits

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.587460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.186039Z digest=sha256:d94a58041c41f4f5e6ce6cdf06e643b7a5a15044ac591310bad1b730c54eb913

Observation 5f065543-cae1-4658-a75d-66bf879e6adf · outbound

This paper cites URLhttps://aclanthology.org/2024.findings-naacl.229/.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs URLhttps://aclanthology.org/2024.findings-naacl.229/

Reference 19

Resolution
malformed identifier
no resolver link, observed 2026-08-14T04:18:07.214750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.214750Z digest=sha256:6f3f6168bdc8a92a4a7a983f0eac16a3162c93fd0fa9e2566737a6ced86c1811

Observation 368143bb-c301-43a6-96f4-1519c277ccdc · outbound

This paper cites an unresolved cited work.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:18:09.469404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.223364Z digest=sha256:4ba2f921c194cdb62abbe6a9480913d982145bf8f91632c206181558de257fc4

Observation 8aa8e873-aeb4-409c-8cd7-21377d555d8b · outbound

This paper cites Qwen Team.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Qwen Team

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.406526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.231328Z digest=sha256:c07f900a32610a0f85547f1d6d7cebcfa959cca1cf905dff0f400498b040b6b9

Observation 3d6bb34a-2c46-4c8a-8d39-3e5595bf4294 · outbound

This paper cites DeepSeek-AI.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs DeepSeek-AI

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.341136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.249330Z digest=sha256:d7b3d455f50427901fd6054095caa89bf1d157258024f429eb7f66ae75cada19

Observation f026ab4a-d16c-4f84-9b91-b675ac11f5f2 · outbound

This paper cites Google DeepMind.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Google DeepMind

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.236384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.255437Z digest=sha256:f10989d73fe578f61cb1642b1c2e64e27cfd243d18205a41805f615787d72f5a

Observation a8f7a612-f0fd-4e71-a953-bb34be7ba50a · outbound

This paper cites an unresolved cited work.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:18:09.104749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.261202Z digest=sha256:62495150705a1078d62442d858a8d1ff094c494b137a6ff455db0ed0e0b076b5

Observation f26e52b9-39c8-4f0e-8d05-9a286aaab7e4 · outbound

This paper cites Qiqiang Lin, Muning Wen, Qiuying Peng, Guanyu Nie, Junwei Liao, Jun Wang, Xiaoyun Mo, Jiamu Zhou, Cheng Cheng, Yin Zhao, et al.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Qiqiang Lin, Muning Wen, Qiuying Peng, Guanyu Nie, Junwei Liao, Jun Wang, Xiaoyun Mo, Jiamu Zhou, Cheng Cheng, Yin Zhao, et al

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.268543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.268543Z digest=sha256:7a87020d609916291db977d523f6dee03425de5bad766ee89e4e1bc3031842ed

Observation 14769fae-8a0c-44f1-8d5c-1eb8751057c8 · outbound

This paper cites Accessed: 2026-05-25.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Accessed: 2026-05-25

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.026654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.287222Z digest=sha256:decee7b20c18e6cea557a694f28ddf8645fcde0dd576cd341d04b8045887f9b6

Observation c47f70df-02f7-498f-823e-b26079cea3a4 · outbound

This paper cites co/Team-ACE/ToolACE-2.5-Llama-3.1-8B.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs co/Team-ACE/ToolACE-2.5-Llama-3.1-8B

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.944279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.296085Z digest=sha256:85a665dd1f44ec10a11e91f7255953430b89fd827c048e7fc7bf9b75916a048a

Observation 105f1510-6014-43e5-a1a1-c0022f71cd55 · outbound

This paper cites Accessed: 2026-05-25.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Accessed: 2026-05-25

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.869221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.319485Z digest=sha256:fee2849ee4ae12bdbd4897e1f307dfe8e672665d0c25d8804a9ca89b14e750be

Observation 99a4a60f-b52c-4512-8e2c-3f3a6665f622 · outbound

This paper cites Direct Multi-Turn Preference Optimization for Language Agents.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Direct Multi-Turn Preference Optimization for Language Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.327158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.327158Z digest=sha256:976ee950db1a9dffbe0d879313b19ed191ea86c35db1e07d312b018f0bc93e35

Observation 06c927ac-55a2-43a5-84b3-8487cd956336 · outbound

This paper cites URL https://aclanthology.org/2026.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs URL https://aclanthology.org/2026

Reference 30

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.720918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.359764Z digest=sha256:6f8879548b2c2d17da7d597245fa70113c30166a2a9e47b8ebe3af5b86e5eee1

Observation 32b744bc-020d-46c1-8c25-765f86d01d3a · outbound

This paper cites URL https://aclanthology.org/2026.findings-acl.1080/.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs URL https://aclanthology.org/2026.findings-acl.1080/

Reference 31

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.689459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.395956Z digest=sha256:6d284843644ae21dedde58aacba6012881e4183895cb6b3ba50c434f6afc3756

Observation 51e3e151-5882-4d0f-bbe5-9ab7da35cd4c · outbound

This paper cites 36Kr GreenLeaf retail rising star.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs 36Kr GreenLeaf retail rising star

Reference 32

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.595153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.427066Z digest=sha256:abf5127d8f4e6e9eabc3299f6be4e80c2509d1b39e0994b97f142881c58ad76d

Observation 72c0a1b2-46ad-4482-844b-ef89a96ff2d5 · outbound

This paper cites B.3 Quantitative Persona Diversity Audit We complement the field-coverage analysis with a quantitative audit of the ten selected profiles.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs B.3 Quantitative Persona Diversity Audit We complement the field-coverage analysis with a quantitative audit of the ten selected profiles

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.832172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.458135Z digest=sha256:be36674c8f5ccd873bc4a8191239d821c9533b553077e793ba43b39ffb39b2d0

Observation 07fe565e-04c8-4585-a2c9-2d88c98ca057 · outbound

This paper cites Left: token Jaccard similarity.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Left: token Jaccard similarity

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.777001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:07.485591Z digest=sha256:67b24bb41388f3f2fa1550addbea51c548ab3d84a532511e8c0361123548a783

Observation 0f9fa615-9808-413b-88f4-21cf915a913d · outbound

This paper cites doi:10.18653/v1/2023.emnlp-main.187.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs doi:10.18653/v1/2023.emnlp-main.187

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.818994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.818994Z digest=sha256:76fb5a214a5b4712c60be00050bea8394733d027d97e9740a4b1ef695ddc2f6e

Observation 2f9c32b5-d519-4eb7-9de1-e782eedbb5e4 · outbound

This paper cites Know me, respond to me: Benchmarking llms for dynamic user profiling and personalized responses at scale.arXiv preprint arXiv:2504.14225,.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Know me, respond to me: Benchmarking llms for dynamic user profiling and personalized responses at scale.arXiv preprint arXiv:2504.14225,

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.605781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.605781Z digest=sha256:f2cce70ec204f954985bc1225affebb1e014cac9de7a416156208ce1d0931ca8

Observation 27ae0a1f-ea3c-470c-9f0b-b490c0c9ee14 · outbound

This paper cites Personalens: A benchmark for personalization evaluation in conversational ai assistants.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Personalens: A benchmark for personalization evaluation in conversational ai assistants

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:10.014822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:06.644752Z digest=sha256:e3110424fe7124f292dbe1ff8d1f662345057c21b75fbde022b507d2294afc16

Observation 1ef11c0d-b1c0-41ea-bc04-2227aea61bed · outbound

This paper cites Petoolllm: Towards personalized tool learning in large language models.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Petoolllm: Towards personalized tool learning in large language models

Reference 2026

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.802674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:18:06.948152Z digest=sha256:ffe7e4ec0b7b2ec079c3431a3c45f5e7aebe05d99c45dc31054a5b887ccbc8ca

Pith citing papers

No inbound Pith citation observations are available.