Pith. sign in

Paper Citation Record · LEDGER

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2608.09507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09507 v2

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:20:54.562259Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ae82675b-e238-4162-b92d-5cdf0a56ff94 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.224787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.224787Z digest=sha256:69973a7499854b765f930d0d37ef49a7ebf41c20b54b7e33df9e3b55696398a3

Observation 39f9f91b-623f-4a7c-a01e-e8a0273c1f07 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.232418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.232418Z digest=sha256:a994f3118aaaefdde1fd2a62a130199976b6d9bf418f1e443f2b9230900a9948

Observation 026823ba-e51a-43fc-bc84-62f34b641006 · outbound

This paper cites an unresolved cited work.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.238370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.238370Z digest=sha256:34d0dcd1fe19a7783fce796556605a431f9a842383c277f7917dc3fca76fc8e7

Observation 24d3e69d-80f8-4cd7-8ae0-8fb14b9cc38b · outbound

This paper cites Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.243621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.243621Z digest=sha256:9e4e24724db9b8321318f66417dc5508a3d7efa980af18c734c51082eda85bd3

Observation 18a3f310-a064-415e-868c-23eb8c2273d9 · outbound

This paper cites Persona-.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Persona-

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.250331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.250331Z digest=sha256:782f93097b632493a2aa3f1a41493a2847900338e37a333e2607544ffc75d418

Observation 9480197e-2490-41bd-a1f7-c3f7a39ca38a · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.255594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.255594Z digest=sha256:827eb2caaa89a9754fd535209092d8b44556cbce9dfdbd7a732765b94022b3c7

Observation 5b2ed914-b6c6-4578-97d6-3cbdb622b41f · outbound

This paper cites Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Retrieval Augmented Generation with Collaborative Filtering for Personalized Text Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.260314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.260314Z digest=sha256:b29218452336de24815b781c4ffcff2af36399eed1edac482978b7a7eed4c1af

Observation 320b4752-14ac-4892-81c5-d47968e78b6d · outbound

This paper cites Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.265392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.265392Z digest=sha256:c0a6f02e7bb1f141532a9221c98ef1a926c79a60672bcf78c2869e7192b71eda

Observation 153e08dc-fac5-4ba5-8673-13989f4fa870 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.270445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.270445Z digest=sha256:356bd2805bfacf0e01e77eb1a091f0f11c034f2aad389f0490e3e50f9f034742

Observation be97fe2b-dc40-4a0b-8871-5d5920a27ef5 · outbound

This paper cites ACM Transactions on Information Systems , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ACM Transactions on Information Systems , volume =

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.276499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.276499Z digest=sha256:f645378ecb5f9015477519a6fb25a9cb32d7542206a9f8f899fc334b78727605

Observation f7d1e735-bc66-45f3-a473-bde7620b331c · outbound

This paper cites 2025 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , url =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.281254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.281254Z digest=sha256:d2b40d722fb9e4ba44b5936d4fa476eab005f238a4d92cf28498299b52995d69

Observation 64625d7a-b7aa-4152-8c50-a10eeb4cb6f9 · outbound

This paper cites Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Personalized Soups: Personalized Large Language Model Alignment via Post-hoc Parameter Merging

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.286992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.286992Z digest=sha256:794e1fa9a16f4e5bf43306d1795a89aa15ea84e02c149f09b72755f247a2853f

Observation 88f85b19-4c63-4c7b-924d-41198c3dc9b5 · outbound

This paper cites 2024 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , note =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.292501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.292501Z digest=sha256:92b7ea66837854e6aa3b7a12e0a34d7f870222fb652018ed3b4ad8c81f29d805

Observation 62c67ef1-a70c-44ee-9e2c-99037d1620dd · outbound

This paper cites Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Integrating Summarization and Retrieval for Enhanced Personalization via Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.297753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.297753Z digest=sha256:2986dcd3e56b950c7c8dc963c9125b10ff4e599c3a901e386b03ae5cb60b023b

Observation 2debc7b9-0f3f-4960-8d30-e20482b33f80 · outbound

This paper cites Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 1st Workshop on Customizable NLP: Progress and Challenges in Customizing NLP for a Domain, Application, Group, or Individual (CustomNLP4U) , pages =

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.303176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.303176Z digest=sha256:a869215451c79f3cd5c862c7ae54ea6e9d1e810308dfab6f553ca3c136592730

Observation c1b9c427-574e-4bc3-8145-1b51167d1b64 · outbound

This paper cites arXiv preprint arXiv:2601.04963 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2601.04963 , year =

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.308080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.308080Z digest=sha256:f0e09180aebe6f654c186f62ac6e26d0f8292410cf8b3ff163b9c0cbd31cd1eb

Observation 003d1eae-dd1f-413e-a607-e01a16c5d69f · outbound

This paper cites International Conference on Learning Representations (ICLR) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Learning Representations (ICLR) , year =

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.313613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.313613Z digest=sha256:456f1954787bf709bfd908e9a8329707eda872e52b92718259b57ca38743933c

Observation bca672d5-3904-42a6-a9c2-7ad5296deba1 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.319100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.319100Z digest=sha256:ec5439b36ec04a991112927487a96488422d7bac60479c4d9e4a04e9ba1d9295

Observation e8d4f5aa-5b6c-4986-84cc-b8a4e1a78b91 · outbound

This paper cites 2024 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , eprint =

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.323577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.323577Z digest=sha256:502897a833d7a1e3fc6e4da4973932e1106645c1dfc5f4b64be2e3a46d3a3027

Observation c99fa9db-4093-4c82-8a09-40d16ee88f6f · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , pages =

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.327979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.327979Z digest=sha256:8c35fa02230aa447290bc1385e0893425a55c420af9c50d259702190fdff2db2

Observation 4836f9bb-6b98-4bae-9c22-988a7b7250b3 · outbound

This paper cites and Stoica, Ion and Gonzalez, Joseph E.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning and Stoica, Ion and Gonzalez, Joseph E

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.332553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.332553Z digest=sha256:209ca20d646af1089db8273ac52ad02d04cb59a282c6400784ff6e750633d066

Observation cae205b1-dfa2-4d1b-8977-66bca40fe932 · outbound

This paper cites 2024 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , doi =

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.337451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.337451Z digest=sha256:82de095a231140052fc6635087f229e02c6cf41f4b1e7a22e8be1be26f9122a2

Observation 817c7abc-8062-45bf-a596-1df3785895b5 · outbound

This paper cites 2025 , note =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , note =

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.342651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.342651Z digest=sha256:fb0751035ae300ff1ae6867b123f7181f9872151fc86b01582ef1c6dee13dcaf

Observation 9f8056a1-8615-4807-8055-dd4402e6930e · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.347138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.347138Z digest=sha256:d7933386c30ca1eb3592384fedc533e2abfbeb25f04add0a837aeca657653c2d

Observation 156dbdaf-87e3-4ece-a776-357c83a40e79 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.352017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.352017Z digest=sha256:eae63e502c183c7bfc08e5080a23b6fb85ce5a5270b2625aa90cfb560b997319

Observation e34a18c8-23a8-416b-8efa-1d8c21feb82a · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.358114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.358114Z digest=sha256:7fef49eaf80a114ec2a2f743e222bd2ed6adfdf8cdd50557fcbe6d344870763c

Observation f837fcaa-535e-4c40-a4c2-1fea889e2d1c · outbound

This paper cites International Conference on Machine Learning (ICML) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning International Conference on Machine Learning (ICML) , pages =

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.364604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.364604Z digest=sha256:963375dd305311a5cad956121cdc5445a43cdd9e58fc699114a563647349cd57

Observation 88c55ad9-3583-450f-b56a-e9a957a5ac58 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.369511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.369511Z digest=sha256:6f5e71fef8b23e07d129eceed43126fe5a5bee2e94bfb42564c940de861c6ac0

Observation 4fac8e80-6c29-4383-8a68-2ff079a9582a · outbound

This paper cites 2026 , month = feb, howpublished =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , month = feb, howpublished =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.373808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.373808Z digest=sha256:61b572c4d5c12f4882bdd774063f534b17dbd8fdfec9d3cd9859aade0c1772ed

Observation fd017266-f053-4a27-b598-2124c60ba001 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.378310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.378310Z digest=sha256:9b3e07fe1da7bc225c5bdc1394fdeee47314fcb6ba1ae3a12748347c78df097a

Observation 002250e1-3403-40d5-a64f-1557d93ab6b3 · outbound

This paper cites The Llama 3 Herd of Models.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning The Llama 3 Herd of Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.383505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.383505Z digest=sha256:0a778b1b175038d18fb1bc793cc1d0b7ab3581462ab395831ab0b6041bcc973e

Observation db907c52-e4f9-444f-8789-5f5c849feb01 · outbound

This paper cites 2023 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , url =

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.399636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.399636Z digest=sha256:257042fbdebf9bf8e5034da714e2799ed6285d5bea0ce342fe5de19f4aae166c

Observation c97964e3-e829-4a56-90ed-246d48c2a753 · outbound

This paper cites 2024 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , url =

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.405937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.405937Z digest=sha256:d4f5c2c72ea83da682c637b970ba687f938eb1b74bbb90dae2206bbe708e0f10

Observation 5f08a450-3df1-4e1d-a64b-51a39d969642 · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: ACL 2024 , pages =

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.412309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.412309Z digest=sha256:626cef24f2d67e406af87dec5e7545a03c516f5defeba42ccb730de63bd2bfb2

Observation 53911f4e-2d67-474d-bf9f-706751177e7a · outbound

This paper cites Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing (EMNLP) , pages =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.418138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.418138Z digest=sha256:cee8bb43998acf38dc5e286a9494a559dad58316a0451737d48dfc20e41edb32

Observation 460022c5-43af-4052-a44b-d60a3ebf5828 · outbound

This paper cites Transactions of the Association for Computational Linguistics , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Transactions of the Association for Computational Linguistics , year =

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.423506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.423506Z digest=sha256:3c710bbf09b50615c9ccf2a8ae2bd317240917393c52b95d2a325762facba8af

Observation 63fee2e0-b204-41fe-add0-0c285345669c · outbound

This paper cites Training language models to follow instructions with human feedback , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Training language models to follow instructions with human feedback , url =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.428777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.428777Z digest=sha256:828d1d0a8d6b8a7cd2e257fff8db76c75bf75ff25a32e604ee0bfee9606a2560

Observation 4edbd32c-8642-49a8-9d42-5d7a18f03594 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.433416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.433416Z digest=sha256:a88f3d78ecc631fd4b17fe53a717cce3c314ae9b9b3217e64851e90b255f7370

Observation c195d0b6-4e2d-42f6-8f8c-28652830f166 · outbound

This paper cites 2026 , url =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , url =

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.438783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.438783Z digest=sha256:fb871304a7e35de72e8fd71750704157e32af884a98ff57e0b90eebee5cd88e3

Observation 05a189b1-8b41-4f15-b58a-35ca5e0f119f · outbound

This paper cites Proceedings of the ACM Web Conference 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the ACM Web Conference 2024 , pages =

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.443255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.443255Z digest=sha256:f2ae9f88bce4b010531a0434897ee1f7c0a5b230a6fa50adba234ae63d5dc403

Observation bfc3c5bb-add7-4554-9e47-743807865343 · outbound

This paper cites Recommendation as Language Processing (.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Recommendation as Language Processing (

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.447799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.447799Z digest=sha256:00e2d22b07d8104ec2a8393f255f71ffa7cfc3d63b2615a5f4caf0c67fb0c34c

Observation e96c8d95-5928-44e5-af39-e3458ab69c6a · outbound

This paper cites 2023 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , doi =

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.453108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.453108Z digest=sha256:9457e0acf692dcc4364e5dc34e9f3cc77ae9caae89ff2ec3e20f21c2a0397f48

Observation f301ccb2-91ac-492a-b361-7d08736e47eb · outbound

This paper cites 2024 , publisher =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2024 , publisher =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.459138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.459138Z digest=sha256:4fcf0f3e6101a972455581024f78d0aa3f11863c3d6891d8f949c438841b5f05

Observation 82f0fc42-0dc4-47c6-97dd-88e6bf3d9d1c · outbound

This paper cites User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning User-Specific Dialogue Generation with User Profile-Aware Pre-Training Model and Parameter-Efficient Fine-Tuning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.465098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.465098Z digest=sha256:d4f260b9a84fc3bb2c59a3f89da7eb0f82839e9d31bde149c343f1a3ca63c2f1

Observation 5a19cf73-f598-4091-a234-74ed9fac0e8a · outbound

This paper cites 2025 , doi =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , doi =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.471041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.471041Z digest=sha256:0e49f35c157a813ead04b308c4bf7f67bc8484e29848b82df0d940d86830ec16

Observation 6f810578-554f-49ab-8670-1630a31e0c3a · outbound

This paper cites Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.477081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.477081Z digest=sha256:282ac4cb6f5cc950703c8d05e928cd351c0640a0708a717194bd8f1095429acd

Observation a4211d96-7bf3-4b26-bb04-5c51ad00d0da · outbound

This paper cites Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Optimization Methods for Personalizing Large Language Models through Retrieval Augmentation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.482148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.482148Z digest=sha256:02d0549dd5671202eab8ca9974cdbe5568d744180e09a9d23495232f09f26109

Observation a909d038-0674-4f6b-a339-a60562909258 · outbound

This paper cites arXiv preprint arXiv:2507.13579 , year =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2507.13579 , year =

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.487135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.487135Z digest=sha256:89db39dcacfbff8693427124388efb007dc30f1815da6d446679b9a3ede80ee5

Observation fa996684-2315-4430-a87a-5d24ff0af2ce · outbound

This paper cites Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Findings of the Association for Computational Linguistics: EMNLP 2024 , pages =

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.491550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.491550Z digest=sha256:284763262e1cfad525f7d20cbe4a90e5ea17af12f49574ca1d8c2c671fbdf2e2

Observation ce8361b2-7509-4a4b-864f-a83d416df745 · outbound

This paper cites 2025 , eprint =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2025 , eprint =

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.496529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.496529Z digest=sha256:480dfa4297d9eddb381063e0daa795874cee05bbf95dd97c6b5410c18523ae6b

Observation 1a3beb4d-455f-45ef-a9df-b8e9b6903d96 · outbound

This paper cites Computer , volume =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Computer , volume =

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.501666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.501666Z digest=sha256:50e9dc6eee822fc83e997952b402ee4bf56e5d3c50d377be610fe7b968368788

Observation d642d4e2-91a8-430e-9f34-79f7304e2fb1 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.506403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.506403Z digest=sha256:50fd81ba5023eea8c70e87f52730bae3f36bd03a5b82e013b577d8c544411707

Observation c09f0f08-5bae-4f2d-8468-5144845ee513 · outbound

This paper cites 2026 , eprint=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2026 , eprint=

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.514162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.514162Z digest=sha256:9a651536a6dafcf6b0a60e187fd1e0675cde9f210da6db4e09dfd3db6ae4aa56

Observation 4991e912-469b-485f-b7ff-b930ebbc2cdf · outbound

This paper cites Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.520361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.520361Z digest=sha256:bb6640681bd888032fdb98c7fecbadcb84f7ef71603fc059b747451e4525bcac

Observation 015d8e13-d63c-4f2c-a6fb-46c1b23a7f8d · outbound

This paper cites MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.525842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.525842Z digest=sha256:eb6d5540cbbecdac7cf5e1a0cf50e7cd9f2fef9cec0e75d2e2aa52dfc3177d5c

Observation 94e6d47b-0d01-4fdc-9997-b350606b5fe2 · outbound

This paper cites 2023 , booktitle =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2023 , booktitle =

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.530953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.530953Z digest=sha256:2e603b6b86277a2351af43a0ca372d5282f21fdd677561e731ddee8ebc720203

Observation f999b699-0ade-4e85-8b9a-975f49990e83 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , pages =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning Proceedings of the 41st International Conference on Machine Learning , pages =

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.537265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.537265Z digest=sha256:a61635f70574e745018d51125fac5611df021ce399eb73b3d1c7ee48239e4311

Observation 8c6e6f33-0ad3-47a4-9854-6ce618d21324 · outbound

This paper cites ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ``In-Dialogues We Learn'': Towards Personalized Dialogue Without Pre-defined Profiles through In-Dialogue Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.542639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.542639Z digest=sha256:7ab90de45de9817382b5e21568a8bdd519a17ad8751288a4c6ff348bba9bd66b

Observation 124cd6a0-43bc-4d5b-a78d-19bb23873811 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.548139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.548139Z digest=sha256:b871cb7732ecda6c464aee78d7b345caa5026189177035310987ef9aeba80be8

Observation 776b781b-648d-41c4-8465-fe3e8f4e16d8 · outbound

This paper cites ArXiv , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning ArXiv , year=

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.552629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.552629Z digest=sha256:06c02374011b6164edc6a0f7b331bbd245a847467f7097c478015545641a1967

Observation e1b129b9-814b-4ea4-97c8-112fa0f93a8f · outbound

This paper cites arXiv preprint arXiv:2603.25973 , year=.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning arXiv preprint arXiv:2603.25973 , year=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.557234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.557234Z digest=sha256:819c067c192918e67a8a152384f55df63f8acc1cf9ab929d969ff31765a5d1cd

Observation 75e39b32-0311-4ce8-b631-3db3ffb090d5 · outbound

This paper cites 2009 , journal =.

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning 2009 , journal =

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-14T04:20:54.562259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:20:54.562259Z digest=sha256:771ab9da5dac4c64c578257b591295bb99662f3361e74b9a59c90b28f741a2cf

Pith citing papers

No inbound Pith citation observations are available.