Pith. sign in

Paper Citation Record · LEDGER

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization

As of 8 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 3 inbound Pith citation observations for arXiv:2506.12915.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12915 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:42:49.542276Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T14:51:27.110839Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:26:22.230033Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact2
  • verified fuzzy18
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation df9fe52e-ff70-4ff4-afbf-c7761ecfa2dc · outbound

This paper cites Llama 3 model card.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Llama 3 model card

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.843270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:43.116952Z digest=sha256:942ccdd79deb74f051d3cf3a59b1600cbfbf31952a2bf3fa968b8867da5b14d4

Observation b19c539b-7d07-4fe1-bbde-cbdbdad264cb · outbound

This paper cites Introducing Claude, 2023.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Introducing Claude, 2023

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.828770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:43.188670Z digest=sha256:f7a82059529148bb2f48e584d36da34536ee7e28a71eb1a8fef3bac2d99d56e3

Observation ff1e70a9-6682-4a2b-a50d-a011166f801d · outbound

This paper cites Explainable recommendations via attentive multi-persona collaborative filtering.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Explainable recommendations via attentive multi-persona collaborative filtering

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.813694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:43.265144Z digest=sha256:71943fdca268f93b89ee8799246aa3b5b7518ee95ab4fb4375a4d37d16978c88

Observation f3b75099-335c-4262-b5c2-1fe692280546 · outbound

This paper cites an unresolved cited work.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:50.796199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:43.356103Z digest=sha256:7f4d74ea768370a5c4c38f13d849e1471421d5f72e30e774ceefcce29b0296e1

Observation 561a77ed-1958-4873-8564-51d59a221274 · outbound

This paper cites Large Language Models for User Interest Journeys.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Large Language Models for User Interest Journeys

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:43.472270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:43.472270Z digest=sha256:d0705f517aec07c15a68dcbe471e0ab8244409a810827b5503a766d6d857cee0

Observation 8a3216b6-73a9-4a1e-9a0d-be5db0003b2b · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:43.629485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:43.629485Z digest=sha256:60d1af1283251ff723fbca442a1c2bdfaaf3b6292284e313e7c500a6e8e4ab6d

Observation d8741c97-87a6-4dd9-acba-06bdd36e5ccf · outbound

This paper cites Language-based user profiles for recommendation, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Language-based user profiles for recommendation, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.780640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:43.726112Z digest=sha256:cdd8450bc00090549258c173e8f9a39e0bbf60c901552c6296072dee83d50bad

Observation 899e24ab-2fc9-4c62-acf9-65dd0abb82dc · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:43.815017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:43.815017Z digest=sha256:7464da0f890a8d8e8ef25f91a3997b39ca8867b920c20351898ae63837963947

Observation d72be3e2-a298-4fb3-8d48-4f70bb024aab · outbound

This paper cites RLHF Workflow: From Reward Modeling to Online RLHF.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization RLHF Workflow: From Reward Modeling to Online RLHF

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:43.921624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:43.921624Z digest=sha256:81362251e863e0a75889e91127ff8fda8ccb9b98fe9f7a6b6d0deb00b5dd6173

Observation 605a7477-2042-4cab-a746-468a1fb2b8d0 · outbound

This paper cites Quantile Regression for Distributional Reward Models in RLHF.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Quantile Regression for Distributional Reward Models in RLHF

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.017650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.017650Z digest=sha256:d03e33c84089759d378f7e558f0c7a7deeadb56f123078568cacfd82d5308252

Observation 45001a41-529d-4669-9d0c-f91bbbfda397 · outbound

This paper cites Statistics (international student edition).Pisani, R.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Statistics (international student edition).Pisani, R

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.765154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:44.098073Z digest=sha256:006f7b9476086d76934fef3c1ccae7780b85e56c65db4f4d8f4ccd2676984061

Observation 03d740ed-0988-4b0f-a89b-85fbd99e33db · outbound

This paper cites ASSISTGUI: Task-Oriented Desktop Graphical User Interface Automation.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization ASSISTGUI: Task-Oriented Desktop Graphical User Interface Automation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.206786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.206786Z digest=sha256:198909be0e408e4cdab5bf0f9a2736602ddae5eeb02f44d8663cbb38df7d5e92

Observation 1193d218-ba07-48e8-89ba-7d602c4bbcfe · outbound

This paper cites Maxwell Harper and Joseph A.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Maxwell Harper and Joseph A

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.309282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.309282Z digest=sha256:dd3f6c26d727259cac5c864ed5620f6a6bdbea2aba66f13326c6bf4772a74744

Observation b1884a62-b91b-4376-a3a7-d49738113ca5 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization LoRA: Low-Rank Adaptation of Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.407417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.407417Z digest=sha256:074a033248e84984156a1f16a947879cd36a636235623ccbe6b26f67d2f6bb64

Observation 1d126f49-7c51-46de-99e0-93dbcd3e385a · outbound

This paper cites Mistral 7B.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Mistral 7B

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.528247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.528247Z digest=sha256:bb62345b371d89d2b5821cfacb84e7e310daf58f1e9656a952bc89aa0b9f51e7

Observation fd144a6d-b9b9-4cfc-9103-81a078686f19 · outbound

This paper cites Know me, respond to me: Benchmarking llms for dynamic user profiling and personalized responses at scale.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Know me, respond to me: Benchmarking llms for dynamic user profiling and personalized responses at scale

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.598568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.598568Z digest=sha256:1123117d4595f04b8df812565387d32612f9a7bb9865b31cd7c50c40e4c39f20

Observation dbe554b2-9e60-4dd1-a25c-75daaa960b0f · outbound

This paper cites Do llms understand user preferences? evaluating llms on user rating prediction, 2023.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Do llms understand user preferences? evaluating llms on user rating prediction, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.749534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:44.661027Z digest=sha256:6afd5d18685ab843386779666f6d7aaca10af3b5ceee7e9ec660d8d850ba2d92

Observation acce208d-69f1-44cb-9511-5df271303cc5 · outbound

This paper cites Smith, and Hannaneh Hajishirzi.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Smith, and Hannaneh Hajishirzi

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.732943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:44.747736Z digest=sha256:a9bbde60e641a87b9b950382be4810da2430de0c100b04d54dc672ccf102629b

Observation 9082a722-79eb-41b9-b5cd-4d17e4b93a5f · outbound

This paper cites Teach LLMs to Personalize -- An Approach inspired by Writing Education.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Teach LLMs to Personalize -- An Approach inspired by Writing Education

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.839306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.839306Z digest=sha256:5bd9db20b436b7456ffd6b927f1e8cb122a4ffea615ead27749af66831fd084f

Observation 0b93c821-47e3-4fe1-8aa4-28568b20c44b · outbound

This paper cites EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:44.993579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:44.993579Z digest=sha256:a1bf422b58f24f46e7e8e50ecc0c1797c632d57d71cae91271fa0559b233c311

Observation 025f6729-5c25-481f-bdfe-384501a9f724 · outbound

This paper cites Personal llm agents: Insights and survey about the capability, efficiency and security, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Personal llm agents: Insights and survey about the capability, efficiency and security, 2024

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:45.206984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:45.206984Z digest=sha256:362c9b11e4ba31c58fed1d7a8a45a47762dbc54c2176403f733129833f58b3a7

Observation 973c623b-a330-4568-a22f-3e7161ae9a37 · outbound

This paper cites DeepSeek-V3 Technical Report.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization DeepSeek-V3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:45.417302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:45.417302Z digest=sha256:13b670f6f2a4e491d0b0b59e2b0df54cba0e9b6ae68aadb2e1d63783279e1560

Observation de613ca4-0ea1-418f-be08-dc47640c140d · outbound

This paper cites Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:45.662556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:45.662556Z digest=sha256:e43545cf39a47fbcb05106012f07fff949b03d2c279b0a3e47ea275b625a2b32

Observation 55e67858-b126-43f9-ac1b-997055aa356e · outbound

This paper cites Gaia: a benchmark for general ai assistants.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Gaia: a benchmark for general ai assistants

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:45.826955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:45.826955Z digest=sha256:83ea38482710fa408bc16f0b20d4e5c21c6f95111b44c76226dc146dfb63537f

Observation 4dabe5ad-e1e6-4a06-b7a2-938676365973 · outbound

This paper cites Inf-orm-llama3.1-70b, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Inf-orm-llama3.1-70b, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.695554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:45.986279Z digest=sha256:ddbfcb20219e02b038f05aac1ceb67afdd942008b851f13c3335970ed2668016

Observation 9d26339a-a4ce-4341-8da2-a6ddfd21d52d · outbound

This paper cites Pearl: Personalizing Large Language Model Writing Assistants with Generation-Calibrated Retrievers.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Pearl: Personalizing Large Language Model Writing Assistants with Generation-Calibrated Retrievers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:46.217896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:46.217896Z digest=sha256:9e4b9fa8af7855304687d11da12fadc54c1e5b33bdb138791c1411a7faca6d96

Observation a502846c-f74d-497d-bbef-3624da7b143c · outbound

This paper cites Justifying recommendations using distantly-labeled reviews and fine-grained aspects.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Justifying recommendations using distantly-labeled reviews and fine-grained aspects

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:46.468846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:46.468846Z digest=sha256:4b396d83991e1fad0011a84e7456ae48dc43e8168281cf29d9b31b20d60e57fc

Observation d1461068-c92c-4a21-ac22-b1f37d22cb80 · outbound

This paper cites an unresolved cited work.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:46.669467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:46.669467Z digest=sha256:ff53323dc32b933b743f7191e8d695fac938cc49a372a9dc2d76616f920f24b0

Observation f960dc81-d99d-495c-91f7-ec993c0142c3 · outbound

This paper cites Offsetbias: Leveraging debiased data for tuning evaluators, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Offsetbias: Leveraging debiased data for tuning evaluators, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.669232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:46.759553Z digest=sha256:e05213850ef08adabf16ec13d2610473e8077e97c04cfbe213b3f454939d03c0

Observation f83332af-f9fb-4ed0-8352-46c1c07a1adc · outbound

This paper cites Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:46.830983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:46.830983Z digest=sha256:784ea1a2d1db1614ebc7f1443f6f4f0c6101a84776c35f9bf058322cfbb21687

Observation e299f189-bbd1-4955-8105-9907101978a1 · outbound

This paper cites Optimization methods for personalizing large language models through retrieval augmentation.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Optimization methods for personalizing large language models through retrieval augmentation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.653364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:46.893593Z digest=sha256:4010b3355f84e75b781b8aaf27e9d7ae56cb6b69a5b226d2a5bdfbdbf98c3923

Observation 38b19d84-5506-4f08-969a-747f6f9600da · outbound

This paper cites Lamp: When large language models meet personalization, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Lamp: When large language models meet personalization, 2024

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.637116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:46.993342Z digest=sha256:6265f9c88947edc5331ea3d90ef95997da46d3f4a8b2f291b6e5a91e91fdedef

Observation 7bdf8faf-4cd0-42c0-a1b0-dbfbb3010824 · outbound

This paper cites User Modeling in the Era of Large Language Models: Current Research and Future Directions.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization User Modeling in the Era of Large Language Models: Current Research and Future Directions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.082342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.082342Z digest=sha256:6e24f1ed85e2db64a65ce27c18f3332acbd5165ff36cbc4aa9388517a463e2cb

Observation d461b08c-aee2-4354-9315-d2dac31d9b72 · outbound

This paper cites Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.145964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.145964Z digest=sha256:815d0b92010a9702ce6196f48d3dc38564627da24da750d1a614523ac4521972

Observation 434d6938-67d5-464d-9f87-cec2fd64484b · outbound

This paper cites RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:42:50.086889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:47.232242Z digest=sha256:c4225e8af7690a85ee5568a89439d4bbdcabeb6d2708484f658d366facf0229a

Observation 0685e883-6cb0-4d03-b974-b7dbd8af8505 · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, March 2025.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Qwq-32b: Embracing the power of reinforcement learning, March 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.309915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.309915Z digest=sha256:8ad92d825d30ba826facae7010cbc44e45878c99b22f393e18654454c42b32c0

Observation 52a40200-fdec-478b-9b9b-eb2ab9254003 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.380532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.380532Z digest=sha256:eecbc79e42631d5655c810625c70bc3d43b902c19be439a901d997dc49f78da1

Observation eb62ce9f-152e-48f4-bb12-f0185e72ab63 · outbound

This paper cites Interpretable preferences via multi- objective reward modeling and mixture-of-experts.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Interpretable preferences via multi- objective reward modeling and mixture-of-experts

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.488090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.488090Z digest=sha256:8eeb291a07919d61a77f1b1cd65643426a6849c8c0a722c16b59417f60d2b11b

Observation 6a019766-1fae-4bd3-be1d-cea19a4539d4 · outbound

This paper cites Weaver: Foundation Models for Creative Writing.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Weaver: Foundation Models for Creative Writing

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.548157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.548157Z digest=sha256:7be6525556d4f752299731a62dc90c1635abd9c7069274e706f9c368a3c24b60

Observation 07794934-94f4-48a7-ab18-1e99d21aad81 · outbound

This paper cites AI PERSONA: Towards Life-long Personalization of LLMs.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization AI PERSONA: Towards Life-long Personalization of LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.668122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.668122Z digest=sha256:0bda67df7de06ea4c9fc9a0f68a3911304979d01b65c54a1cefa0ff9ac222fc0

Observation 8d710010-a9d1-489c-94c5-ad874f411ade · outbound

This paper cites HelpSteer: Multi-attribute Helpfulness Dataset for SteerLM.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization HelpSteer: Multi-attribute Helpfulness Dataset for SteerLM

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.768903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.768903Z digest=sha256:d63c0f7bd5ca596d76365273d3720349fdcaef19eee7b26465a0e71964f28d5b

Observation b3b23bd3-640a-4e84-89ff-091dd43c11d9 · outbound

This paper cites HelpSteer2: Open-source dataset for training top-performing reward models.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization HelpSteer2: Open-source dataset for training top-performing reward models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:47.880878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:47.880878Z digest=sha256:8473078c8d5ec6b243b654d5236615ffaf2283292dbcba0563a9585dda454d93

Observation aa31f723-4050-45ba-beab-792317fd412c · outbound

This paper cites Personalized large language models, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Personalized large language models, 2024

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.601764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:47.977115Z digest=sha256:f71b221c597732608ac58450c61dc2cd35978347429581c92373e2aef6cdfcc6

Observation c2f88625-db39-4e3e-8f3f-1633a63acba6 · outbound

This paper cites Travelplanner: A benchmark for real-world planning with language agents.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Travelplanner: A benchmark for real-world planning with language agents

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.586415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:48.038867Z digest=sha256:b1938122c280d7289b8c3931cf2eebd15067a215804f23651d4d6b9919b61b7a

Observation 443bcc92-0eec-47b2-91f2-6e23338da4aa · outbound

This paper cites Iterative preference learning from human feedback: Bridging theory and practice for rlhf under kl-constraint, 2024.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Iterative preference learning from human feedback: Bridging theory and practice for rlhf under kl-constraint, 2024

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:48.114443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:48.114443Z digest=sha256:9df8e6eef842cdca3fffe9291bea307ab3dac156234b93c3dc77f5edaa397c0d

Observation 90af4112-5601-42c0-8992-089a64fef636 · outbound

This paper cites Qwen2.5 Technical Report.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Qwen2.5 Technical Report

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:48.319868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:48.319868Z digest=sha256:35a99899a0dd9adb51b60ce93ddbc05f8dd3b128aace6ab27e3dd0c127d1e69d

Observation 48cd27ac-c3ea-4593-8bf8-e05e9953898a · outbound

This paper cites RefGPT: Dialogue Generation of GPT, by GPT, and for GPT.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization RefGPT: Dialogue Generation of GPT, by GPT, and for GPT

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:42:49.854512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:48.395777Z digest=sha256:5fc27e4279d7a948b2ae37d61646ea38d653ba94be70a3f03dfb7adb5f0dfd92

Observation fd8a4f3d-b645-4051-a713-0c8618677e9a · outbound

This paper cites Palr: Personalization aware llms for recommendation, 2023.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Palr: Personalization aware llms for recommendation, 2023

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.559366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:48.483560Z digest=sha256:574ecd399cf1b355b4d97542b74ee2714af932b904f7100e07ad1ed491a237e8

Observation c8ca9a81-b836-42c8-b79c-0ed0a04d842c · outbound

This paper cites BookGPT: A General Framework for Book Recommendation Empowered by Large Language Model.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization BookGPT: A General Framework for Book Recommendation Empowered by Large Language Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:48.564482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:48.564482Z digest=sha256:6ca102c6170e5ab14f88ce267b3995a222c3c9b785bced58890813e99cae67b5

Observation 51d0cc83-eb94-4122-8a29-3233366d4fdf · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Instruction-Following Evaluation for Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:48.631231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:48.631231Z digest=sha256:a8acd6d0c2201e69a2925e34157fd5ebe185583b3bf2b548af09190dddca723d

Observation 8788018d-c404-4c45-bc0a-847d301b557f · outbound

This paper cites Learning to predict persona information for dialogue personalization without explicit persona description.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Learning to predict persona information for dialogue personalization without explicit persona description

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:48.711798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:48.711798Z digest=sha256:d597507f09b78b6395fb697ebd05c17d003315543d04841fee63002f9710a9e2

Observation 94f043c8-85a3-4f6f-9cca-51230c272405 · outbound

This paper cites PersonalLLM: Tailoring LLMs to Individual Preferences.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization PersonalLLM: Tailoring LLMs to Individual Preferences

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:48.799048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:42:48.799048Z digest=sha256:9fa9fa400fd13f9bb7b81d9631928edc0a6fa86d2491054e648e36cb099b7356

Observation aa98592d-827f-404c-b1a4-3de0e13ce052 · outbound

This paper cites Avoid purely factual queries.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Avoid purely factual queries

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.544671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:48.864789Z digest=sha256:6e2f78ea300eb76396ab1f72fbe86d8e627d0c14c1ef1fde3593a0b2f1545531

Observation cdbe10b5-461b-4df8-bd19-a0bf147a9a1d · outbound

This paper cites Do not awkwardly mash memory scenes and persona configuration.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Do not awkwardly mash memory scenes and persona configuration

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.530177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:48.938392Z digest=sha256:39962cefa1417973ea10f4db5fbeca4d4ed6e6293fa70aa7ba07ea58c7413b2b

Observation f0962368-c738-45f3-801e-ccc15cbcfc52 · outbound

This paper cites an unresolved cited work.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:50.514814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:49.000542Z digest=sha256:7df1aba29a00a557843c893ba95ea0a235f1a86b7a0ad18b4984e1c286370577

Observation 08bba89e-7836-4ab4-b1ed-d1d2f60846c5 · outbound

This paper cites an unresolved cited work.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:50.496231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:49.072393Z digest=sha256:3ba543c660895e5d76393ff4b6c7e198d9b2652dc5807092ba8d28a7e90a96fc

Observation 6f0cf8fc-877e-4f07-8129-e9174fd3afa4 · outbound

This paper cites Instead, infer what the user might want based on their existing persona configuration, rather than just combining field details.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Instead, infer what the user might want based on their existing persona configuration, rather than just combining field details

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.481803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:49.172406Z digest=sha256:5a23556397a06faa7d842600d54ba6e7836c7dc5aa9dbf5b8e01cfa8e1942b4f

Observation a89e8741-4e83-4f7e-992d-b71feb342576 · outbound

This paper cites an unresolved cited work.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:50.466452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:49.319156Z digest=sha256:0750bcdfdbfc5951dee50daadd18de990cf9ffb34f032af9f6c444172f5bb842

Observation d39cdb81-053f-4132-bcff-6707cbe688bc · outbound

This paper cites an unresolved cited work.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:42:50.451650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:49.423886Z digest=sha256:e144e94ba86af6f7f4a1e4c9befe72481bc1ce9cb1ce9e6e63316e6f196244f9

Observation 862f2935-76d3-4266-85f9-654c411cc260 · outbound

This paper cites For instance, add overly general or irrelevant information, or provide vague or overly broad advice.

PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization For instance, add overly general or irrelevant information, or provide vague or overly broad advice

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:42:50.435777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:42:49.542276Z digest=sha256:96a73c61d8a75b54540cf892a154f01f4870f5e1b675266a38ffc0a9fdcb5194

Pith citing papers

Observation 1beb8346-8819-4887-8d18-355c24811f45 · inbound

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment cites this paper.

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:51:08.344417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T14:50:23.587894Z digest=sha256:e6d1dcbba70fc9450c3eeed8de6ee03f86cf84e5c60fd640f15f8a6a4cf7393f

Observation 1a8b8716-b32f-4be0-85d8-05f974e0118d · inbound

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization cites this paper.

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:20.508725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:51:27.110839Z digest=sha256:e1737ed0ae75de181afcdf4d9f03e4ff627105e9211bc51fd5b1ac28143c87cd

Observation a79d1622-dd64-44c9-9ceb-3422a5b9d438 · inbound

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces cites this paper.

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces PersonaFeedback: A Large-scale Human-annotated Benchmark For Personalization

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:26:22.232555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T14:25:04.965228Z digest=sha256:41d2840abeca4fa5ed900aaa352b0a5a44a685aa30fbf79f7c6d7e3ccc2892ed