Response-Aware User Memory Selection for LLM Personalization

· 2026 · cs.AI · arXiv 2604.14473

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

open full Pith review browse 1 citing papers arXiv PDF

abstract

A common approach to personalization in large language models (LLMs) is to incorporate a subset of the user memory into the prompt at inference time to guide the model's generation. Existing methods select these subsets primarily using similarity between user memory items and input queries, ignoring how features actually affect the model's response distribution. We propose Response-Utility optimization for Memory Selection (RUMS), a novel method that selects user memory items by measuring the mutual information between a subset of memory and the model's outputs, identifying items that reduce response uncertainty and sharpen predictions beyond semantic similarity. We demonstrate that this information-theoretic foundation enables more principled user memory selection that aligns more closely with human selection compared to state-of-the-art methods, and models $400\times$ larger. Additionally, we show that memory items selected using RUMS result in better response quality compared to existing approaches, while having up to $95\%$ reduction in computational cost.

citation-role summary

background 1

citation-polarity summary

background 1

representative citing papers

When Are LLM Inferences Acceptable? User Reactions and Control Preferences for Inferred Personal Information

cs.HC · 2026-05-11 · unverdicted · novelty 7.0

Users show curiosity over concern toward LLM inferences of personal information, with acceptability depending on context, alignment with expectations, and who uses the inferences rather than just the content.

citing papers explorer

Showing 1 of 1 citing paper.

When Are LLM Inferences Acceptable? User Reactions and Control Preferences for Inferred Personal Information cs.HC · 2026-05-11 · unverdicted · none · ref 14 · internal anchor
Users show curiosity over concern toward LLM inferences of personal information, with acceptability depending on context, alignment with expectations, and who uses the inferences rather than just the content.

Response-Aware User Memory Selection for LLM Personalization

citation-role summary

citation-polarity summary

fields

years

verdicts

roles

polarities

representative citing papers

citing papers explorer