Pith. sign in

Paper Citation Record · LEDGER

Personalized Language Modeling from Personalized Human Feedback

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2402.05133.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.05133 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:08:23.633018Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:57.654015Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a228e9a9-5995-497d-9177-5e7d69c62d2b · inbound

Test-Time Alignment via Hypothesis Reweighting cites this paper.

Test-Time Alignment via Hypothesis Reweighting Personalized Language Modeling from Personalized Human Feedback

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:57:40.443797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-23T06:55:54.051821Z digest=sha256:5c495b5d38982bd015eb0607f13467ec282c779d56156cb1c5702891fc503b5a

Observation 38151b5c-effe-4c85-a42b-0fce1a8e4579 · inbound

Configurable Preference Tuning with Rubric-Guided Synthetic Data cites this paper.

Configurable Preference Tuning with Rubric-Guided Synthetic Data Personalized Language Modeling from Personalized Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:23.633018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:08:23.633018Z digest=sha256:ce13334ad2870d3b16250dfba29f30d01e9b74f08d0cb7390b6b2d1d54b39253

Observation 5cd7afb4-fe5b-495e-9338-84609153062c · inbound

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities cites this paper.

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Personalized Language Modeling from Personalized Human Feedback

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:25.079612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:25.079612Z digest=sha256:09c32ae730d63d9a1617bce24f5b162aececba06284fe4092687c46b4fd6f591

Observation e8175cae-ddf0-4bda-b3dd-1bc248f379dd · inbound

Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions cites this paper.

Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions Personalized Language Modeling from Personalized Human Feedback

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:59.930987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:29:59.930987Z digest=sha256:ca150247d0fd03386d7612610d089bd3ae8da6bd41f9c1e301c7e91b0d789f0d

Observation 8b8a602c-498a-4b8f-b136-85d3e0668285 · inbound

Learning the Value Systems of Societies from Preferences cites this paper.

Learning the Value Systems of Societies from Preferences Personalized Language Modeling from Personalized Human Feedback

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T13:26:21.239176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:26:21.239176Z digest=sha256:f86c61e0fe4ad38930b94624fbe2de3d3fec1c72c8f1afc4509e6734024de5a4

Observation 3cc3cd93-4237-4282-aadf-fdff52fda1b5 · inbound

PREF: Reference-Free Evaluation of Personalised Text Generation in LLMs cites this paper.

PREF: Reference-Free Evaluation of Personalised Text Generation in LLMs Personalized Language Modeling from Personalized Human Feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T22:50:43.010695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:50:43.010695Z digest=sha256:bad2da1c629af18a1ad81ca68576b475fc31b3679012e15555ec2348317f74d5

Observation f2e82074-bd53-4e10-870b-2be899b8e1cd · inbound

T-POP: Test-Time Personalization with Online Preference Feedback cites this paper.

T-POP: Test-Time Personalization with Online Preference Feedback Personalized Language Modeling from Personalized Human Feedback

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:11.492730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:11.492730Z digest=sha256:8fa1a834ca9b1169252c20b9e9c2da9aaf3afdc2d9c88245cff8affb759ba64b

Observation f9183074-5824-4fb1-a3a1-9e8c78344105 · inbound

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference cites this paper.

POPI: Personalizing LLMs via Optimized Natural Language Preference Inference Personalized Language Modeling from Personalized Human Feedback

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:42:24.394040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T05:41:58.231139Z digest=sha256:2cdfed9b646f03a772aedd7e287574535f236972542520efd154e8d5249c9a33

Observation 1fee4a06-e7f3-4d75-a4b1-4865469298c8 · inbound

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs cites this paper.

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs Personalized Language Modeling from Personalized Human Feedback

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T12:42:29.904718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:42:29.904718Z digest=sha256:41c7ea57b54b47ea3d2ec2e8e71e889d2b68f4d9d1db0802fa0fc846c69542ea

Observation 4c320c8b-3ff4-45d1-b69e-6d37ce263bed · inbound

Behavior Latticing: Inferring User Motivations from Unstructured Interactions cites this paper.

Behavior Latticing: Inferring User Motivations from Unstructured Interactions Personalized Language Modeling from Personalized Human Feedback

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:26:03.260928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:10:19.647811Z digest=sha256:0aab69c1817a6d30f564df0493df06c07536cb9dfa02aebeb94f679c54915a85

Observation 72027b06-9be5-4556-84e7-3f79c6f8df40 · inbound

Efficient Personalization of Generative User Interfaces cites this paper.

Efficient Personalization of Generative User Interfaces Personalized Language Modeling from Personalized Human Feedback

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:00:57.711247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:52:34.799158Z digest=sha256:9ba7dc1d4c2a1d6eb25d5a0b35b1b4282893b93243eb5501d9a63b5ce9ec6bc7

Observation 99c2279c-0a30-4afe-9522-1742a333f8db · inbound

Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization cites this paper.

Mobile GUI Agent Privacy Personalization with Trajectory Induced Preference Optimization Personalized Language Modeling from Personalized Human Feedback

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:11:01.200631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:11:49.149334Z digest=sha256:b2ad9cfb911d3ac04187fff9b57a5b32eb95b93427962fd14229fc67d92badba

Observation 90297f47-1cc3-4327-8cf8-b8bf49328478 · inbound

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies cites this paper.

Separable Expert Architecture: Toward Privacy-Preserving LLM Personalization via Composable Adapters and Deletable User Proxies Personalized Language Modeling from Personalized Human Feedback

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T22:34:07.449414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T22:26:34.487400Z digest=sha256:b99a5efe8896833c9f5d55f320238a2731a8b30b71649847bb0cb2bf278dbd01

Observation f2f7e3f0-05d1-42e0-887d-4a05ae2dcbc2 · inbound

CLIPer: Tailoring Diverse User Preference via Classifier-Guided Inference-Time Personalization cites this paper.

CLIPer: Tailoring Diverse User Preference via Classifier-Guided Inference-Time Personalization Personalized Language Modeling from Personalized Human Feedback

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:54.307061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T02:16:28.593349Z digest=sha256:5c504c11f357889c3c771e2aaaee7cf9ab9bda2731d4b0174ad6cdb4841c3dbd

Observation 8e00764b-a21e-4ace-9ef5-7a0157ab1f57 · inbound

Active Query Synthesis for Preference Learning cites this paper.

Active Query Synthesis for Preference Learning Personalized Language Modeling from Personalized Human Feedback

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:23:59.906280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T22:22:57.678463Z digest=sha256:4551860c0001166f2c3609b7a100119ecc85554590743e74ad61086e44cd1fef

Observation 0e157fa7-9075-429f-a80f-b601dbf92094 · inbound

Recon: Reconstruction-Guided Reasoning Synthesis for User Modeling cites this paper.

Recon: Reconstruction-Guided Reasoning Synthesis for User Modeling Personalized Language Modeling from Personalized Human Feedback

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:53:51.214710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T18:50:35.279152Z digest=sha256:008a2cc220b313c88ff701ec9b294492dee0209c4c69681be082339195c885e0

Observation 773b80f2-346f-44d1-a4cf-8de5eea2b131 · inbound

On the Road to Personalized Code Intelligence: Portraiting and Assisting Developers Based on Their In-IDE Behaviors cites this paper.

On the Road to Personalized Code Intelligence: Portraiting and Assisting Developers Based on Their In-IDE Behaviors Personalized Language Modeling from Personalized Human Feedback

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T09:13:16.595573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T06:54:47.741975Z digest=sha256:c18666309ddbe6b4776b51e3174f282e8dfb9c2747d169029702d85b0848400c

Observation 1d81c672-f164-42a3-b185-c5d3804f0b82 · inbound

From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users cites this paper.

From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users Personalized Language Modeling from Personalized Human Feedback

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:42:35.906477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T18:55:19.520180Z digest=sha256:05b2cd99ce0de3264f73f731f4115f1f843ab9b88b725174dbc5524a56f4c409

Observation 8519c2ae-5dbf-4a3f-b070-a12f3190292c · inbound

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization cites this paper.

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization Personalized Language Modeling from Personalized Human Feedback

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:56:20.574467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:51:27.110839Z digest=sha256:01561f7184d488716eab85f3c93c562247fe0f89501bf83ce201b6ddbdb24ebd

Observation 12593219-a43b-412f-a331-ef4198696965 · inbound

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling cites this paper.

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling Personalized Language Modeling from Personalized Human Feedback

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:57:23.624710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:00:05.900814Z digest=sha256:e01090c60987b571157ae0be7fd1f1859a2fa7189f450b68ac54560089112525

Observation b2276f88-1d9a-4c49-9e47-0195534c788e · inbound

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs cites this paper.

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs Personalized Language Modeling from Personalized Human Feedback

Reference 167

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T01:07:30.219986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T16:49:14.243931Z digest=sha256:0e60a0163cb3e54a9be63e6dbfd422728994707c86d80417c40bdbf6e704cc32

Observation 8b0cd9d3-b3b0-4258-87ec-af8ffe28c28b · inbound

Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games cites this paper.

Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games Personalized Language Modeling from Personalized Human Feedback

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:58:57.655330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T01:03:49.101568Z digest=sha256:457bb5d42b446e14afef8eaa0a26fc87f5231c64fbb3ebc4c7754cc8ca7eaef5

Observation 4a6cb344-7fe2-4d1f-a587-93fccda64185 · inbound

CoPersona: Collaborative Persona Graphs for Robust LLM Personalization cites this paper.

CoPersona: Collaborative Persona Graphs for Robust LLM Personalization Personalized Language Modeling from Personalized Human Feedback

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:28:48.429614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T18:20:53.930801Z digest=sha256:e8b5dfc934c86daa444ad9b8981c8782ae79e091aab2c3996c409129c2f92004

Observation b6e96cc9-331d-44ce-86ac-95cbdb09680c · inbound

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex cites this paper.

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex Personalized Language Modeling from Personalized Human Feedback

Reference 80

Resolution
unresolved
no resolver link, observed 2026-07-31T23:51:56.023066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:51:56.023066Z digest=sha256:451d57df0626c2bd4a8cb8ac4bbfc4722c8ff01b0a1a66a906ca3d1a479ebbf6