Pith. sign in

Paper Citation Record · LEDGER

On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2411.02306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.02306 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:07:29.471740Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f988d7e4-0356-480d-96c4-767364624d4e · inbound

The Lock-in Hypothesis: Stagnation by Algorithm cites this paper.

The Lock-in Hypothesis: Stagnation by Algorithm On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:29.471740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:07:29.471740Z digest=sha256:558d5372c0b134ace96ce66e6248be37a8eb8be8051b8ff7ad2d762eccc99766

Observation 0c2aa3db-74cd-4152-8fa1-dfcef47d83d0 · inbound

Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework cites this paper.

Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:38.443628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:38.443628Z digest=sha256:6a3c4a45ca03dcc216802a34525679dd1490a4e395b821dc50b653048eb72544

Observation 1a9a12ca-456d-490e-a71b-77b9dc424b9f · inbound

Mitigating LLM biases toward spurious social contexts using direct preference optimization cites this paper.

Mitigating LLM biases toward spurious social contexts using direct preference optimization On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:33:14.711288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:33:04.433907Z digest=sha256:4cb48e218b1d471aa1a52f4dfdb6e66234f2338d0bb5f2cc88b663204dc32963

Observation e7853648-c97f-4152-b45e-93f4915b0e0a · inbound

Quantifying the Utility of User Simulators for Building Collaborative LLM Assistants cites this paper.

Quantifying the Utility of User Simulators for Building Collaborative LLM Assistants On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:36:47.497987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T02:28:13.317630Z digest=sha256:9f2683f81233dfe4397022c3fef65960309d61f2cbd10cacabcd045b6db726c1

Observation 6f416061-052a-4a36-ad3b-73db504dbbe3 · inbound

Structure from Strategic Interaction & Uncertainty: Risk Sensitive Games for Robust Preference Learning cites this paper.

Structure from Strategic Interaction & Uncertainty: Risk Sensitive Games for Robust Preference Learning On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:26:26.433528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T04:15:24.919355Z digest=sha256:fa7cbff86d6ab6047e38fdf77c84ccc6acc419645a18633ede63494b1dec44c2

Observation 34fee9e2-f2cf-4589-9372-4ee2e14d56ed · inbound

Structure from Strategic Interaction & Uncertainty: Risk Sensitive Games for Robust Preference Learning cites this paper.

Structure from Strategic Interaction & Uncertainty: Risk Sensitive Games for Robust Preference Learning On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:08:05.074380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T22:03:05.102274Z digest=sha256:a2ad2c5fda175d1529949893a29fcabd945da1810478cee0976ba369eb02b361

Observation 3895b28b-088f-46cf-9173-a7d42d07ca6f · inbound

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs cites this paper.

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:41:22.314555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-22T09:38:04.387777Z digest=sha256:85f36903ef74dd6e3c0940c8f9c50bf1b3676a38972af3fa469d3f812892b9da

Observation e0525972-eb1f-4cbe-b8cf-92f9b5659cf7 · inbound

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs cites this paper.

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:41:21.438931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-22T09:38:04.387777Z digest=sha256:db86225bb7f6140d1b5c72f9252d9ed6ec4bd64b2c7fc3784ee6dda2bc59fbff

Observation 32ce9421-a677-49fb-bb53-5592957b5e3a · inbound

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs cites this paper.

Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:34:58.065439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T17:17:39.899234Z digest=sha256:77b9e4893530b472c18e5a82dc822d1eaed071c9ce68984a871eb5e67653be87

Observation 5af3ff9c-5371-481e-aade-f6878d2db672 · inbound

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing cites this paper.

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 131

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.587162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T06:17:01.173495Z digest=sha256:1761dcadb3a1597ed0f690119465c1e066137e6ace8bd2c9920755165310ff1f

Observation 05b8766f-9d6d-4619-8d5b-3c0d94259f68 · inbound

Against Proxy Optimization cites this paper.

Against Proxy Optimization On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:49:46.629443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T08:26:50.595370Z digest=sha256:904cd7f71f33a4c768fe90b1ac00c47c2561364a563531060827083f7f9c71f2

Observation b8249767-90a5-44c7-bcc2-52b47f825175 · inbound

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action cites this paper.

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:44.649060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T05:40:54.002702Z digest=sha256:f55cd8d460644ca2079f1e3d995cdb7c16e009d11b5ecc84da3309404ebebf5e

Observation 948a15f7-17d7-4f76-8b0f-5b311c1ea7e5 · inbound

Psychological Influences of Conversational AI: Research and Design Directions for Reducing Harm and Promoting Well-Being cites this paper.

Psychological Influences of Conversational AI: Research and Design Directions for Reducing Harm and Promoting Well-Being On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback

Reference 158

Resolution
unresolved
no resolver link, observed 2026-07-31T02:25:21.866266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:25:21.866266Z digest=sha256:afe754854dade6c0adbb8425fe5fdd0b6b584f2809d3dc59daeb2de3a3895eac