Pith. sign in

Paper Citation Record · LEDGER

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2608.09507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09507 v2

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:20:54.562259Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae82675b-e238-4162-b92d-5cdf0a56ff94 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.224787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.224787Z digest=sha256:90ef3f6f78377047adefd2ba2a0051977ffc1e78fd4d604e982c359cc9a03a29

Observation 39f9f91b-623f-4a7c-a01e-e8a0273c1f07 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.232418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.232418Z digest=sha256:90119f2f588c4830805442de00d1ea6c347de41408b377873283ddbd12f3ddee

Observation 026823ba-e51a-43fc-bc84-62f34b641006 · outbound

This paper cites an unresolved cited work.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.238370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.238370Z digest=sha256:325cdd1bbd339858076a4b3c9d443967418413261c369af490b2da2a920fba3d

Observation 24d3e69d-80f8-4cd7-8ae0-8fb14b9cc38b · outbound

This paper cites Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.243621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.243621Z digest=sha256:4e678824495704eb3fc355f224044c32f7e146ceb3abbac4abb21fb988910d1f

Observation 18a3f310-a064-415e-868c-23eb8c2273d9 · outbound

This paper cites Persona-.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Persona-

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.250331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.250331Z digest=sha256:fde924d591a2bccce81233701af16a4578c9e04989eabcb15f9fa5ec98f3ba77

Observation 9480197e-2490-41bd-a1f7-c3f7a39ca38a · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.255594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.255594Z digest=sha256:4bf128a67c867f3690e3ff706dfc6a133d270a7d07f9757544b64b2d5c31a578

Observation 5b2ed914-b6c6-4578-97d6-3cbdb622b41f · outbound

This paper cites Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.260314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.260314Z digest=sha256:0653d02b861518c768b53691f1bc6a9aca42a00ba91312b8febc61bd9b623db5

Observation 320b4752-14ac-4892-81c5-d47968e78b6d · outbound

This paper cites Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.265392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.265392Z digest=sha256:ca563cefd42270b0c797daf86c806de030de8b1248fd69b6061a860b7dba8d6f

Observation 153e08dc-fac5-4ba5-8673-13989f4fa870 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.270445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.270445Z digest=sha256:c3f426b7697130a18bf61ca80937ed2fb043b2dc7891c4b4d68893187ea7dc83

Observation be97fe2b-dc40-4a0b-8871-5d5920a27ef5 · outbound

This paper cites ACM Transactions on Information Systems , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ACM Transactions on Information Systems , volume =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.276499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.276499Z digest=sha256:3f460ecbbfcf219d70a07f9be06a59357db1a2ac0072810d12c9397965e62278

Observation f7d1e735-bc66-45f3-a473-bde7620b331c · outbound

This paper cites 2025 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , url =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.281254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.281254Z digest=sha256:20f41c9df58c3f9859415da3d22ec6edb9dc5d17c3b95b755540a32acee4e2bf

Observation 64625d7a-b7aa-4152-8c50-a10eeb4cb6f9 · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.286992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.286992Z digest=sha256:e459ee85ca06510f06fe3fae98d93a63c5a1f7552b7a4b37d2720f2172be1530

Observation 88f85b19-4c63-4c7b-924d-41198c3dc9b5 · outbound

This paper cites 2024 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , note =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.292501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.292501Z digest=sha256:aa51606f9be5f1369932e759a2dc8f10c772895d50e683f827c8b7a7e3dae4a8

Observation 62c67ef1-a70c-44ee-9e2c-99037d1620dd · outbound

This paper cites Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.297753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.297753Z digest=sha256:75f76cb7cdd0468054674fc9015fc31d19bf720ebc6e96328213a25e17bfd265

Observation 2debc7b9-0f3f-4960-8d30-e20482b33f80 · outbound

This paper cites Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.303176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.303176Z digest=sha256:54d3c87d74345638d9dfb7fc8e2760c290b2beda6650a407f20763d420ee8857

Observation c1b9c427-574e-4bc3-8145-1b51167d1b64 · outbound

This paper cites arXiv preprint arXiv:2601.04963 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2601.04963 , year =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.308080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.308080Z digest=sha256:11956a9b7d595327191fb9790b46ebd5d60591b3dd0fa588653e99bb808945a7

Observation 003d1eae-dd1f-413e-a607-e01a16c5d69f · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.313613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.313613Z digest=sha256:27db92ea8487f7c7b770f3e27e2c0de4b7369e721404fae739e04006bfdadb04

Observation bca672d5-3904-42a6-a9c2-7ad5296deba1 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.319100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.319100Z digest=sha256:364e595429eb5a7a5bd5e502ca9b8df824cd5079306cdc57f14b274fc3097508

Observation e8d4f5aa-5b6c-4986-84cc-b8a4e1a78b91 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.323577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.323577Z digest=sha256:094f77bc5dee367b12dc273343022c8f2ca797fa52ba20ffd78a0b5e3c4883b9

Observation c99fa9db-4093-4c82-8a09-40d16ee88f6f · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , pages =

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.327979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.327979Z digest=sha256:0d370866fca43582f1e7cbf68305e2e96b1ea9dbf99c273627eb26b461895b39

Observation 4836f9bb-6b98-4bae-9c22-988a7b7250b3 · outbound

This paper cites and Stoica, Ion and Gonzalez, Joseph E.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning and Stoica, Ion and Gonzalez, Joseph E

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.332553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.332553Z digest=sha256:1a68fe25451591db313f8434a07080cb63e0d22dd55a1ebf07dad88e402a8580

Observation cae205b1-dfa2-4d1b-8977-66bca40fe932 · outbound

This paper cites 2024 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , doi =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.337451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.337451Z digest=sha256:76515b6fc76d870bb904f2601e3a0dc053305d2ec05f69cf360e56290d56accf

Observation 817c7abc-8062-45bf-a596-1df3785895b5 · outbound

This paper cites 2025 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , note =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.342651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.342651Z digest=sha256:b2e760281e7cbf48dd8df6384d116abc13fa6e1a357408a74d73e9f2df5956a4

Observation 9f8056a1-8615-4807-8055-dd4402e6930e · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.347138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.347138Z digest=sha256:8ee02d9ce3e363655aaa5be85edf7f89f678d0831961005aa62b40f5b32db397

Observation 156dbdaf-87e3-4ece-a776-357c83a40e79 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.352017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.352017Z digest=sha256:16115376d4b633e07c258f78c6a9f02b40efa8d9c4513bc9bf7ae78eb47939f6

Observation e34a18c8-23a8-416b-8efa-1d8c21feb82a · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.358114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.358114Z digest=sha256:f751021395cbc435c5d20f3d6e86724a607d724a981a15bd35f84571432acee0

Observation f837fcaa-535e-4c40-a4c2-1fea889e2d1c · outbound

This paper cites International Conference on Machine Learning (ICML) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Machine Learning (ICML) , pages =

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.364604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.364604Z digest=sha256:f9135ddf4a0dd526fc35455eed29c3d86ba903c4dc6220471ae0910664fd17cc

Observation 88c55ad9-3583-450f-b56a-e9a957a5ac58 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.369511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.369511Z digest=sha256:33b2561b2df5736be4c1e0c8653b6bf18f5ab2a4fb2c0c7ff95af2839362f8d8

Observation 4fac8e80-6c29-4383-8a68-2ff079a9582a · outbound

This paper cites 2026 , month = feb, howpublished =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , month = feb, howpublished =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.373808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.373808Z digest=sha256:48b303d3276f4c6a1a84f23d5aa8ebbb3401e5380acf3e500a99aec09ca97b70

Observation fd017266-f053-4a27-b598-2124c60ba001 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.378310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.378310Z digest=sha256:320703da9628bf0307db65c40a069f76c6edb680a19baffe5156c38100f56624

Observation 002250e1-3403-40d5-a64f-1557d93ab6b3 · outbound

This paper cites The Llama 3 Herd of Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning The Llama 3 Herd of Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.383505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.383505Z digest=sha256:ccf3dcd7f889badbf5681032f36ff0fbc89c80bdb93c10f4cde81ebba0273274

Observation db907c52-e4f9-444f-8789-5f5c849feb01 · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.399636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.399636Z digest=sha256:e613fb46d0b43b72a67d0ba5fa1a262e39d5055f617db64ebb97cffddcdcc132

Observation c97964e3-e829-4a56-90ed-246d48c2a753 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.405937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.405937Z digest=sha256:38f96fd50f165121c253e9cbfb70b6e0573f832daeeac25c406b5836b340203f

Observation 5f08a450-3df1-4e1d-a64b-51a39d969642 · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: ACL 2024 , pages =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.412309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.412309Z digest=sha256:0fbd1ff5464126ae091944fef2f0c968aa5f83888ffb025410837623d1d82ae4

Observation 53911f4e-2d67-474d-bf9f-706751177e7a · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.418138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.418138Z digest=sha256:e92b1c6c629165c49dbbef9ef9fbe7cc680ff3de724dfc3e560036cec748a628

Observation 460022c5-43af-4052-a44b-d60a3ebf5828 · outbound

This paper cites Transactions of the Association for Computational Linguistics , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Transactions of the Association for Computational Linguistics , year =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.423506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.423506Z digest=sha256:7522df0c03d3fa84b46d0cfff074e0dcf9aea63b42b5875ae407affd1443be1e

Observation 63fee2e0-b204-41fe-add0-0c285345669c · outbound

This paper cites Training language models to follow instructions with human feedback , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Training language models to follow instructions with human feedback , url =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.428777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.428777Z digest=sha256:8385762c923138d94175043a73f1a98431ddf8ba969b478e92561908555e13e3

Observation 4edbd32c-8642-49a8-9d42-5d7a18f03594 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.433416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.433416Z digest=sha256:2db0a50a31a8cb34e1cbc8c1c9d5be30fb0e4c792fa19b40878067d4a120a2a8

Observation c195d0b6-4e2d-42f6-8f8c-28652830f166 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.438783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.438783Z digest=sha256:a1ffa3c299272fd757636dbf7a10fbb44943a91f8266e0ad10b731b2bdab1e82

Observation 05a189b1-8b41-4f15-b58a-35ca5e0f119f · outbound

This paper cites Proceedings of the ACM Web Conference 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the ACM Web Conference 2024 , pages =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.443255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.443255Z digest=sha256:cf1d6a5c54ee6dde8a34fa74dcb78ea6f73f0755325d44f32a2d1c7153bf5f3f

Observation bfc3c5bb-add7-4554-9e47-743807865343 · outbound

This paper cites Recommendation as Language Processing (.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Recommendation as Language Processing (

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.447799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.447799Z digest=sha256:c05ec27a2244791fe313efd2d80d1691d655fadae0b78b43b6c5ed68cf994716

Observation e96c8d95-5928-44e5-af39-e3458ab69c6a · outbound

This paper cites 2023 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , doi =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.453108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.453108Z digest=sha256:62cc315ab1e6c28cd7e44272e65de3798764c26856e17f7fdb54184dca084820

Observation f301ccb2-91ac-492a-b361-7d08736e47eb · outbound

This paper cites 2024 , publisher =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , publisher =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.459138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.459138Z digest=sha256:bbc9437d0c3cf58382b2b94b24e1d955e9ac2f736c09032a2058cc9a151945b3

Observation 82f0fc42-0dc4-47c6-97dd-88e6bf3d9d1c · outbound

This paper cites User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.465098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.465098Z digest=sha256:d0297c74f079e756a647a4f212c9ef798fe942a16c25fc73b60f567b08595e5e

Observation 5a19cf73-f598-4091-a234-74ed9fac0e8a · outbound

This paper cites 2025 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , doi =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.471041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.471041Z digest=sha256:ffdf342d542cd0555799ed753f7fb5994303d8652829e19b69444d328b22e804

Observation 6f810578-554f-49ab-8670-1630a31e0c3a · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.477081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.477081Z digest=sha256:8f54b7d01ae182d36e72292d72513273c7be6a2db1dd98166bb14be246f1867c

Observation a4211d96-7bf3-4b26-bb04-5c51ad00d0da · outbound

This paper cites Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.482148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.482148Z digest=sha256:988a60429791b3169fe3504fd59b17d214a95f3440de22133ec1b6c9c5efda5f

Observation a909d038-0674-4f6b-a339-a60562909258 · outbound

This paper cites arXiv preprint arXiv:2507.13579 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2507.13579 , year =

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.487135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.487135Z digest=sha256:5e6a5073e65faa845916232b33741089509b5352d63c52aab642d5437a4a2951

Observation fa996684-2315-4430-a87a-5d24ff0af2ce · outbound

This paper cites Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.491550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.491550Z digest=sha256:fd844be105727022040502ac18a2a6f5d170556410bc1e341f0e17dba9d94937

Observation ce8361b2-7509-4a4b-864f-a83d416df745 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.496529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.496529Z digest=sha256:0e1785bbb480e768c6c3c7e7a53d0ff9c19a889762b6195b3a97d680ffeb3ee6

Observation 1a3beb4d-455f-45ef-a9df-b8e9b6903d96 · outbound

This paper cites Computer , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Computer , volume =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.501666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.501666Z digest=sha256:a36fa6329b320e6cdd27ceca0dfeff6f4deb2f4339a2309cfebd6dfe9f24a17f

Observation d642d4e2-91a8-430e-9f34-79f7304e2fb1 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.506403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.506403Z digest=sha256:3393753da399b3b9c4df3713a6252a8b934cbe9ba5cc5f6a0abcba3ca9c996e7

Observation c09f0f08-5bae-4f2d-8468-5144845ee513 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.514162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.514162Z digest=sha256:4be9aef81161ee7e59113713508143bc157619e360452091e8e0466800dbfb2b

Observation 4991e912-469b-485f-b7ff-b930ebbc2cdf · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.520361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.520361Z digest=sha256:ef5f5012975522570e728aff49544904be56ffdcdda0c8cc1906fd28207b354b

Observation 015d8e13-d63c-4f2c-a6fb-46c1b23a7f8d · outbound

This paper cites MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.525842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.525842Z digest=sha256:719e9a3af6bdc2cf571ceb0399b4695b4a73803b3ae5c8536bc8121fca163c2a

Observation 94e6d47b-0d01-4fdc-9997-b350606b5fe2 · outbound

This paper cites 2023 , booktitle =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , booktitle =

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.530953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.530953Z digest=sha256:265ebe8958755b43586f0389931de4b888102ef0aaf5058068bcaf350a1cf406

Observation f999b699-0ade-4e85-8b9a-975f49990e83 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 41st International Conference on Machine Learning , pages =

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.537265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.537265Z digest=sha256:678956e6b20f944831f099e35e35e1f290cf456a1f61312a27cc8810e730557b

Observation 8c6e6f33-0ad3-47a4-9854-6ce618d21324 · outbound

This paper cites ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.542639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.542639Z digest=sha256:148b85da5ae81449a28f743c37eff51d7f9672e65235a1fbf45f171cdf94ca44

Observation 124cd6a0-43bc-4d5b-a78d-19bb23873811 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.548139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.548139Z digest=sha256:1fafa37ff41b874052164a3b4e844d2853835c3d4492c0639a95d9a17d85db11

Observation 776b781b-648d-41c4-8465-fe3e8f4e16d8 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.552629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.552629Z digest=sha256:d17d4ca070602edd60e524e5de502961c3bcf01749019351d3d206560488310a

Observation e1b129b9-814b-4ea4-97c8-112fa0f93a8f · outbound

This paper cites arXiv preprint arXiv:2603.25973 , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2603.25973 , year=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.557234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.557234Z digest=sha256:3c38a621d9b8d6ca29fe80254816a41d7a664d5796e38c0bb5040d4e4f67cf10

Observation 75e39b32-0311-4ce8-b631-3db3ffb090d5 · outbound

This paper cites 2009 , journal =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2009 , journal =

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.562259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.562259Z digest=sha256:b78569ebc33424f3f69b18a6e10d63bb1df7b42c0d3c8d570bfb66d84f0d8669

Pith citing papers

No inbound Pith citation observations are available.