Pith. sign in

Paper Citation Record · LEDGER

Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2210.01790.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.01790 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:04:57.439817Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9ccb2324-bc9e-45fe-a20c-4e180e231e83 · inbound

Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs cites this paper.

Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T00:04:57.439817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:04:57.439817Z digest=sha256:5f2b8021e1bec86114a5262326ecd7baf0c3749b460523b683d376d6d2e30e0b

Observation bd00104a-e810-4aa8-8ada-3afa05d52c3c · inbound

Technical Risks of (Lethal) Autonomous Weapons Systems cites this paper.

Technical Risks of (Lethal) Autonomous Weapons Systems Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T19:10:56.735518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:10:56.735518Z digest=sha256:9d64275e3703426aab3ba7947a71e1fe375d933dbd01a35a5f2b70898f5559c1

Observation c01b1197-dea2-4fd7-8e52-628342c9f5b4 · inbound

Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies cites this paper.

Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 154

Resolution
unresolved
no resolver link, observed 2026-08-07T13:05:32.661716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:05:32.661716Z digest=sha256:93d7c88b8fea9e6687ba0992e8966c2050f0fa3a64bfab0188d74c18d039c382

Observation 5239c1e3-0a0c-4092-bdc3-736d05106210 · inbound

AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models cites this paper.

AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 1270

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:13.140632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:45:13.140632Z digest=sha256:7314989777b1c09be775aae88051aac248e41652568a27df96b2c881be63b158

Observation d89b77a1-e7f6-46bb-81c3-f37439138105 · inbound

Mitigating Goal Misgeneralization via Minimax Regret cites this paper.

Mitigating Goal Misgeneralization via Minimax Regret Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.268826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.268826Z digest=sha256:90bced134e5ca956453059b38606b2ec0f9c8436f2abd52bb6828ff7c6ae971b

Observation 3a50fbda-be39-4508-afa8-021011168d71 · inbound

Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language cites this paper.

Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T20:18:15.459547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:18:15.459547Z digest=sha256:a34463c663c0d8cc821a5f99499a2ab1dd10bf203d5f108af229bdd4bfe3d773

Observation b86a2244-c5df-4e12-bc22-fbe0aadbd572 · inbound

Semantic Convergence: Investigating Shared Representations Across Scaled LLMs cites this paper.

Semantic Convergence: Investigating Shared Representations Across Scaled LLMs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:39:46.553247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:39:46.553247Z digest=sha256:7a0b5a3ede27d91e7bf72ee460fcbf7a4f16ec623776332e000dab06f3a61741

Observation 6602d481-ab81-4b73-a57b-572c15993da5 · inbound

Interpretable Electrophysiological Features of Resting-State EEG Capture Cortical Network Dynamics in Parkinsons Disease cites this paper.

Interpretable Electrophysiological Features of Resting-State EEG Capture Cortical Network Dynamics in Parkinsons Disease Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T14:23:49.346727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:23:49.346727Z digest=sha256:829ca2df789c0eea28830b5e23d73f979efe30150df8d54ab078bd826dcd06d8

Observation fc413ef3-77aa-4e5c-8cb2-9ff910d1225d · inbound

Why Does Agentic Safety Fail to Generalize Across Tasks? cites this paper.

Why Does Agentic Safety Fail to Generalize Across Tasks? Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:06:00.148917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T01:55:38.554161Z digest=sha256:4e9762e4f08b78cff77a1514951aa5f388e1bdd77d000a512f1c5824d0eeeca7

Observation 56e0c30d-f7bc-4a2c-b018-96b0a27829c9 · inbound

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training cites this paper.

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:32:24.261475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T06:30:51.812541Z digest=sha256:3d2c072425404c21a7ff5179548220f9f11e43cc44a9f6052ab35ecc67273cf6

Observation d6b62d88-4d53-4aff-9442-31fa43fc5179 · inbound

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack cites this paper.

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:32:56.848195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T20:31:50.043920Z digest=sha256:1211b2f44f3d23076c037be3fee886e40f82491d8c2d85ee91756d6cc5f56b99

Observation 584b3b41-465c-4a39-9201-3e01e7339d36 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 224

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:54.260119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:0700c7f4a2ff461f5e678f3e5cbc7f623d306ed89af8d5924427f3499b674b08

Observation d5631898-2d29-4a4c-a850-41288180f501 · inbound

Understanding Goal Generalisation in Sequential Reinforcement Learning cites this paper.

Understanding Goal Generalisation in Sequential Reinforcement Learning Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:50:20.839480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-25T04:49:50.034743Z digest=sha256:3702af2a138677b9b748b81eb6cae793a98f36aee1eb1a3afbccb5ef9e3df743

Observation 33f4e2df-1111-4edb-8047-70cf3b707736 · inbound

Voluntary Collusion with Secret Tools in Competing LLM Agents cites this paper.

Voluntary Collusion with Secret Tools in Competing LLM Agents Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:40.769914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T17:01:22.732390Z digest=sha256:e02df8eeb9ecb2e61e18baa97372adc10c87f0777430ab9ae5232903a990b702

Observation c543ef0d-d36a-478e-9534-d9046f192b04 · inbound

Beyond Value Benchmarks: Measuring Value-Structure Alignment in Large Language Models via Symmetric Q-Sorts cites this paper.

Beyond Value Benchmarks: Measuring Value-Structure Alignment in Large Language Models via Symmetric Q-Sorts Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.879344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T12:13:12.018506Z digest=sha256:754c9cedb892259ab9169bdd38de1ffae74669895cc52c52d6da50e2e7f45e5d

Observation dd3ba819-e0b0-426e-95b0-13bf0ee1beeb · inbound

A Scalable Approach to Evaluating Moral Sensitivity in LLMs cites this paper.

A Scalable Approach to Evaluating Moral Sensitivity in LLMs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 156

Resolution
unresolved
no resolver link, observed 2026-07-12T05:44:33.099337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T05:44:33.099337Z digest=sha256:e942f9746a68eff21385e848e2992257fdf7892df8f4e51e3b21ffe72c43a837

Observation 140911e7-4a29-46ba-b69b-c6c9fecb38a0 · inbound

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary cites this paper.

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T07:20:39.148585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:20:39.148585Z digest=sha256:f298da2a0a1aaa09e1af8613e97d961f277994cea5b1919b2dbf4a9792efe194

Observation 1f48ad79-a499-40c0-b165-e44d061f11bd · inbound

Innocuous-Seeming Data, Latent Ideology: Ideological Generalisation in Finetuned LLMs cites this paper.

Innocuous-Seeming Data, Latent Ideology: Ideological Generalisation in Finetuned LLMs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T00:51:18.184662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:51:18.184662Z digest=sha256:c924d72d566265b32de85d88228541613546885d35d1a37d9cc2ea248def2a69

Observation c871cfe9-711c-4206-b6f4-fbc093d770e1 · inbound

AI Value Alignment for Evolving Social Norms cites this paper.

AI Value Alignment for Evolving Social Norms Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T15:16:55.934050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:16:55.934050Z digest=sha256:397cd8418d04412dc3d49556c777f3d3afb0526d674d56f222a5900d4b663dcd

Observation 37c95f28-7fa9-4c6e-aacb-8e43d5861797 · inbound

Response drift across frontier large language models cites this paper.

Response drift across frontier large language models Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T13:58:47.006884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:58:47.006884Z digest=sha256:69bca22b53ab1fb6623486bcf4118a74261e0316f23a4fe919548c3432dd87b0

Observation 6712e349-7885-44c6-930a-99684fdaaf2e · inbound

Co-design of LLM-based preference agents: participation may drive overtrust cites this paper.

Co-design of LLM-based preference agents: participation may drive overtrust Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T06:48:46.131233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:48:46.131233Z digest=sha256:ccd62b160a7e9561747f03a9eabd9c2347247f9f28e1db146e4f602b8660f7f2

Observation d0b3bd55-23ba-4f82-ac7f-4e85e3a62dc2 · inbound

Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives cites this paper.

Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:12.902670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T05:36:12.902670Z digest=sha256:74e7066fcf9a28cdeb12e2e84071a6d8ccdb7db1211f6d62afd077f4655c561b