Pith. sign in

Paper Citation Record · LEDGER

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs

As of 9 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2608.05246.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05246 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:17:53.524604Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact2
  • verified fuzzy17
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cbe5b08b-a1a0-4570-84a2-5ea7da5335df · outbound

This paper cites Learning to reason for multi-step retrieval of personal context in personalized question answering.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Learning to reason for multi-step retrieval of personal context in personalized question answering

Reference 1

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T17:17:54.748244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.272051Z digest=sha256:b3c4a50b4988b7f82aad2e84d309beaf1b31991f219f96f1276fb19846991f30

Observation 79ad8bfc-f4b9-429c-8877-4842ba5b5e35 · outbound

This paper cites Large language models empowered personalized web agents.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Large language models empowered personalized web agents

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.277727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.277727Z digest=sha256:809eea365b834a228002ea7aa0e411e2aa7c73811cebe63306d85cf0b494304d

Observation 6b6bc56a-3184-44fe-a44e-f73234dc6201 · outbound

This paper cites Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.283149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.283149Z digest=sha256:a93ce36f8f8cd95abcd744baa553b64f4e4eafd800f18b9ffcb1f295ca9dd8e7

Observation 05ac2751-17c0-4bf0-90b6-984eefa4c23c · outbound

This paper cites KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.289378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.289378Z digest=sha256:99097640d5d83ffd1842c73ec9469ee4413947f0e4945ab69735138ac299e5d0

Observation 647d158d-7195-422d-87bc-e353944c7d53 · outbound

This paper cites POPI: Personalizing LLMs via Optimized Natural Language Preference Inference.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs POPI: Personalizing LLMs via Optimized Natural Language Preference Inference

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.294909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.294909Z digest=sha256:74e4b5455f2951b8d9ad3013586798ab26d0bfa60960d84ce79ad0be15747ba4

Observation 5fa22ae1-bd35-49d5-a584-db57a9fdf609 · outbound

This paper cites Lifebench: A benchmark for long-horizon multi-source memory, 2026.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Lifebench: A benchmark for long-horizon multi-source memory, 2026

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.301349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.301349Z digest=sha256:fec7178d93132c29669b6a3c7d501852bce504b6c3600959f3013422133a55d1

Observation d566f43c-52a3-489b-b7d7-a44398374f6c · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.306961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.306961Z digest=sha256:34fea77a2312f87f7196519c1dc903bc0cbf0cda05829a3afe5ea837fd027aed

Observation 1f13e3a0-abd7-4c66-8078-1ab509693913 · outbound

This paper cites Lifesim: Long-horizon user life simulator for personalized assistant evaluation.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Lifesim: Long-horizon user life simulator for personalized assistant evaluation

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.125871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.312072Z digest=sha256:7f3b124243295629dc1bc96b3f54cd454bae08d8e7e7af991b64c7ee65b3abff

Observation 015c68cd-1c79-4cd7-87f4-2dc8ef03ff31 · outbound

This paper cites A survey on personalized alignment—the missing piece for large language models in real-world applications.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs A survey on personalized alignment—the missing piece for large language models in real-world applications

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.104016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.316730Z digest=sha256:456f0f394e55b015a00f7a5e0b0d3f9b5c057383ca6ce4c1d55cd1467d08fd77

Observation 2f3e4e44-ae26-4184-bf7f-b33dbd09a03a · outbound

This paper cites Towards realistic personalization: Evaluating long-horizon preference following in personalized user-llm interactions.arXiv preprint arXiv:2603.04191, 2026.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Towards realistic personalization: Evaluating long-horizon preference following in personalized user-llm interactions.arXiv preprint arXiv:2603.04191, 2026

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.321693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.321693Z digest=sha256:d77fe7ee100173088182c691a2dbfe43906eaf2d7a38de567eb63468719478d2

Observation 5a833f00-6104-4aff-8872-1a955379b6ea · outbound

This paper cites Computing inter-rater reliability and its variance in the presence of high agreement.British Journal of Mathematical and Statistical Psychology, 61(1):29–48, 2008.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Computing inter-rater reliability and its variance in the presence of high agreement.British Journal of Mathematical and Statistical Psychology, 61(1):29–48, 2008

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.327020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.327020Z digest=sha256:38daae0adac2d28c82d882ce57f6602ed2b7174b11524ec6f070579a0ccdeeb4

Observation 4e65dcc8-0313-48e6-ba2f-143f54d7cea1 · outbound

This paper cites Rap: Retrieval-augmented personal- ization for multimodal large language models.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Rap: Retrieval-augmented personal- ization for multimodal large language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.074174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.331856Z digest=sha256:9b73bec17537a833ff552dceca3bc38ef0bad2a5017b76169b4929fa62f252f2

Observation 56da7c87-ee3c-4df9-b752-e35cd45242dc · outbound

This paper cites Asking the Right Questions: Improving Reasoning with Generated Stepping Stones.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Asking the Right Questions: Improving Reasoning with Generated Stepping Stones

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.336566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.336566Z digest=sha256:1353b94c0cb870ff5c040c430334db650744584784ed9beaf03b9c007adef563

Observation 587f4df2-c4fc-4f95-8a4f-e51722811664 · outbound

This paper cites Op-bench: Benchmarking over-personalization for memory-augmented personalized conversational agents.arXiv preprint arXiv:2601.13722, 2026.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Op-bench: Benchmarking over-personalization for memory-augmented personalized conversational agents.arXiv preprint arXiv:2601.13722, 2026

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.342011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.342011Z digest=sha256:b8d429f190de9193b45905e9d6b266bfb8342c7107b5da03ef62c380213e283c

Observation 26d5f4ea-1045-4c42-a35a-452f625912d9 · outbound

This paper cites Mem-pal: Towards memory-based personalized dialogue assistants for long-term user-agent interaction.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Mem-pal: Towards memory-based personalized dialogue assistants for long-term user-agent interaction

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.056133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.346834Z digest=sha256:2ff1d0503d91ba98f27879f5cd346d6a60e926076845ded037ff9aa707a25586

Observation fc44df1e-9e30-44e3-8dea-6c4ea85bddbd · outbound

This paper cites Taylor, and Dan Roth.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Taylor, and Dan Roth

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.352101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.352101Z digest=sha256:24f8c5fe3eab61364047eb7d5a84b05b71af21dec56702a66f87baabd66b2e29

Observation 09fed038-012f-4d9a-a4e7-d0b7da1e84b5 · outbound

This paper cites Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.356534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.356534Z digest=sha256:7ae5568b5c675f3781a01ff6aacc7b74e361b57b3ec9d905dde9702b9edd49e8

Observation 1761f197-9cf5-44b7-9579-ff4770ea77d5 · outbound

This paper cites Humanllm: Towards personalized understanding and simulation of human nature.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Humanllm: Towards personalized understanding and simulation of human nature

Reference 18

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T17:17:54.163090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.361281Z digest=sha256:30caef2bf4e380b106d1a1f919d56f243d18cdd7770b8c64cd8a2427e293f6d0

Observation 3df11aa5-64a0-4e50-9c98-4ab4a69a979f · outbound

This paper cites Retrieval-augmented genera- tion for knowledge-intensive nlp tasks.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Retrieval-augmented genera- tion for knowledge-intensive nlp tasks

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.039109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.366074Z digest=sha256:d8a7df9874aa7970e671d785c8fdc3ae7ae3e8d395f70f8ad1eb7e0e0e356380

Observation 0ec4ee73-f0bd-48d8-b716-81e5dd09304c · outbound

This paper cites Can llm agents simulate multi-turn human behavior? evidence from real online customer behavior data.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Can llm agents simulate multi-turn human behavior? evidence from real online customer behavior data

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.021118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.370929Z digest=sha256:ccc7aa0fe1a4ae0d011c2cdb30cd1c9deb84f8036013bb72140d2446367763c4

Observation 1f0c53b9-6e91-4f50-94c1-229004c058e5 · outbound

This paper cites Exploring the potential of LLMs as person- alized assistants: Dataset, evaluation, and analysis.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Exploring the potential of LLMs as person- alized assistants: Dataset, evaluation, and analysis

Reference 21

Resolution
verified exact
doi, observed 2026-08-08T17:17:53.595425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.375836Z digest=sha256:a75d155295a91227e272bf5f939fce86de875230832d700fd409a3bfb9bcd057

Observation d126780d-3f88-4567-a0a9-0fda6c6b59fc · outbound

This paper cites Privacybench: A conversational benchmark for evaluating privacy in personalized ai, 2025.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Privacybench: A conversational benchmark for evaluating privacy in personalized ai, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.380675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.380675Z digest=sha256:677775cb8f7498033ef60e7a5856efcc25db82cc6d8747b4e8f7254d4662072c

Observation ad1f9ba3-504a-4ab0-8a64-429516c6ef1e · outbound

This paper cites PersonaVLM: Long-Term Personalized Multimodal LLMs.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs PersonaVLM: Long-Term Personalized Multimodal LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.385274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.385274Z digest=sha256:74fe352f256aff643a38398fd6abc57f69288dd44a88221185ed294942889752

Observation cb7d46fd-ed79-4425-ae21-9c923116454c · outbound

This paper cites On Memory Construction and Retrieval for Personalized Conversational Agents.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs On Memory Construction and Retrieval for Personalized Conversational Agents

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.390227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.390227Z digest=sha256:6c649c996143d713dba30e4497642ffb54a97c8c3f589ad256e751ffa9e259f5

Observation a2e0a7b9-8a7c-4286-a73a-a35006f2af53 · outbound

This paper cites LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.395160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.395160Z digest=sha256:4dd1c2ca8be30eb4b4858a5747cf295104a7fa3d63895493dab342f90268ef59

Observation c4df0f79-a0e7-40ed-aa7c-a6b0fc3768e0 · outbound

This paper cites Lamp-qa: A benchmark for personalized long-form question answering.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Lamp-qa: A benchmark for personalized long-form question answering

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:55.003847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.400822Z digest=sha256:ec2b0dbdd0428b145318bbec4e283444106b66a01b895991cfcbf2caa81dfa4d

Observation b7cb8012-50dc-419e-bf7d-a489e1e3d0e5 · outbound

This paper cites Optimization methods for personalizing large language models through retrieval augmentation.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Optimization methods for personalizing large language models through retrieval augmentation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.405758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.405758Z digest=sha256:896527e56084e836be71da69e4aae03578233a1501435c7f9f15ecbc47c460c2

Observation 636e4262-22b5-469c-be69-20b8decfc0c2 · outbound

This paper cites Lamp: When large language models meet personalization.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Lamp: When large language models meet personalization

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.987973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.410357Z digest=sha256:f30d4a7499751ed1ec0ce1be39616f7af93533f0c2f493bd561a2a89080e4715

Observation 63cdc21d-4945-41e9-b83a-2d50e16f610b · outbound

This paper cites PersonaBench: Evaluating AI models on understanding personal information through accessing (synthetic) private user data.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs PersonaBench: Evaluating AI models on understanding personal information through accessing (synthetic) private user data

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.414867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.414867Z digest=sha256:4593ed6771457ad37b2f580888e1e7ba65d3109d9d1ca14f8d81e0f53733d7cd

Observation ad19aebd-878a-49c0-a267-9cf724497249 · outbound

This paper cites Democratizing large language models via personalized parameter-efficient fine-tuning.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Democratizing large language models via personalized parameter-efficient fine-tuning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.961165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.419312Z digest=sha256:3f2558e9934e201984bfaba8de796cdc4e33f6a069efd6ccf6ee19030e5de2af

Observation 5e34d33a-f1a6-4e94-88bc-38449f3a8511 · outbound

This paper cites In prospect and retrospect: Reflective memory management for long-term personalized dialogue agents.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs In prospect and retrospect: Reflective memory management for long-term personalized dialogue agents

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.945799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.424373Z digest=sha256:6743d08bb50cb9a0f7173f0de30e355d7836289d7491a1938bf1ab49e1c3a3c9

Observation cba440d3-9b69-45e2-a3f3-0aa57e8bb554 · outbound

This paper cites PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.434087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.434087Z digest=sha256:65f47e02e48b1c638c5db350b07a7161754e6e1612cc1bfb02118b0579cf12b3

Observation d24e8727-09ee-4441-ab20-6213f38d17de · outbound

This paper cites OPeRA: A dataset of observation, persona, rationale, and action for evaluating LLMs on human online shopping behavior simulation.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs OPeRA: A dataset of observation, persona, rationale, and action for evaluating LLMs on human online shopping behavior simulation

Reference 33

Resolution
verified exact
doi, observed 2026-08-08T17:17:54.930558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.438950Z digest=sha256:3b035173eee4ecaeba251f054094acba310d00cfd656a8b5feb63175c0393312

Observation d7cabc12-5fd3-4ad5-9e9c-693f2d03cd75 · outbound

This paper cites Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.443925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.443925Z digest=sha256:f451d142daefcfc3cdca98960850ebe8b89a5b978512962e570e28df3edfa53f

Observation 1636cc8f-8755-4d64-a7a7-2d33cc89459b · outbound

This paper cites DynamicMem: A Long-Horizon Memory Benchmark in Real-World Settings.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs DynamicMem: A Long-Horizon Memory Benchmark in Real-World Settings

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.448806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.448806Z digest=sha256:6b593fb79141893bd2f7f4f0d35e44be486de835e3e30c51e446d63c7ca0f56a

Observation 294eab45-e44f-4b36-b70a-debc93e08f27 · outbound

This paper cites Lauvrak, Jon Atle Gulla, and Heri Ramampiaro.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Lauvrak, Jon Atle Gulla, and Heri Ramampiaro

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.913126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.453682Z digest=sha256:0179e4903511aecbad6bf04af2b066264f542fc0f39fdb486cbd7c64bffcd356

Observation d87058c1-ad0c-498c-956d-3ab710d15acf · outbound

This paper cites an unresolved cited work.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.459192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.459192Z digest=sha256:6ac2bee7665e1e0b0e32e4ea48b15d0bc142c2dddb17a6854986e324e0a994c7

Observation 19ce0ba7-0d2c-4154-8a0b-359a578a26b8 · outbound

This paper cites Promax: Exploring the potential of llm-derived profiles with distribution shaping for recommender systems.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Promax: Exploring the potential of llm-derived profiles with distribution shaping for recommender systems

Reference 38

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T17:17:53.739143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.464402Z digest=sha256:121c10983cefad672d9159b9e7b1900fd39d05ee7abc7504b24e0eb03398203b

Observation adffac51-bba2-4278-aec2-65d48b3e81bd · outbound

This paper cites Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.469368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.469368Z digest=sha256:eccfb0e0dcacf4c606041a371af4ea09bd725f5a2c25f82ff3adf3abb009dbbb

Observation d0a33fd7-19bc-4d37-b060-191a6d177e0e · outbound

This paper cites Cohen, and Emine Yilmaz.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Cohen, and Emine Yilmaz

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.474220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.474220Z digest=sha256:bc68a125d976458f2ade739165d09840731d52db44b3642351537e35a5093790

Observation 3bef6ec4-4c44-4604-9663-e53fa6a84e70 · outbound

This paper cites Cantonese cuisine.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Cantonese cuisine

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.479258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.479258Z digest=sha256:f1338d49be7879ca962729c89510d73110e3331c798a6265c9959570ff3ec7b3

Observation 46159ad9-2549-44f9-819d-091c4950fd99 · outbound

This paper cites You tend toward budget-conscious choices.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs You tend toward budget-conscious choices

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.896381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.485499Z digest=sha256:702ed0368b83e6a99628fd69f02f4815b18875854b883e96e5560dec97a22b01

Observation cbacb1d9-02bc-4d7d-911f-7b2fd1fbb502 · outbound

This paper cites Note that cross-domain references serving the response are legitimate; offense occurs only when references are purely demonstrative.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Note that cross-domain references serving the response are legitimate; offense occurs only when references are purely demonstrative

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.880274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.490402Z digest=sha256:3f983b96d777fdd93026600bd1d56be22be1d83aebc75a5e541609bcd26c4102

Observation 91dd12f4-3f53-478f-99a7-0b60caec758d · outbound

This paper cites educate" the user, such as.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs educate" the user, such as

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.863931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.495608Z digest=sha256:c642fc7db590e83480e7249e768aa9292cbfe4ebcb3d436b03248fa80f08b7a3

Observation 8d064032-1750-4f6c-9512-85b3a0f62214 · outbound

This paper cites an unresolved cited work.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:17:54.848030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.500517Z digest=sha256:3c9783c67fb0f4a0c39704c636685f9e5054b9e33ff0f3f323636810dbe93aa2

Observation ad1790ac-5891-4b0f-ad40-1fab75938661 · outbound

This paper cites an unresolved cited work.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:17:54.830944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.505050Z digest=sha256:b3d1dba2018e1fcd24d61d6da434c19fe2de4104eb61180a138acba959205f72

Observation 177d19fa-f24d-4b90-bd51-6c8034977286 · outbound

This paper cites retrieval.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs retrieval

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.815062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.509773Z digest=sha256:fc8755c7b47df9c895c56de9056c9167398f745fef4763350179bcf1ff3a1bd3

Observation c704c8f8-0641-4b77-8a85-e02176aefcae · outbound

This paper cites swap test.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs swap test

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.799387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.515287Z digest=sha256:7c6019db1738880d2f3b6b3d539c4d6d735e5bc4d82c8e96b2f1b2b60c9e4ed5

Observation f44d3304-2982-4065-862d-6d738a01164c · outbound

This paper cites an unresolved cited work.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:17:54.781914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.519873Z digest=sha256:7c3fef4cce938e6d646f36451e5e26a8cb222203931c8275c81fd1520d7731d9

Observation 7f5c301d-97ed-471f-a7b4-055c5b9993f3 · outbound

This paper cites retrieval.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs retrieval

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:17:54.765834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T17:17:53.524604Z digest=sha256:628998124a578fea0d2a0673e9792feec8fb351285a95aa9cb79d601685e4c44

Observation 0e0385c0-3997-4f02-8d7f-73c7f45ff40a · outbound

This paper cites ISBN 979-8-89176-251-0.

LUNAR: Benchmarking Personalized Large Language Models on UNiversal User BehAvioR Logs ISBN 979-8-89176-251-0

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T17:17:53.429103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:17:53.429103Z digest=sha256:9705633436f9dad3f9880c014a3740934e7f2e17d5caa8f5560d0e66636206a0

Pith citing papers

No inbound Pith citation observations are available.