Pith. sign in

Paper Citation Record · LEDGER

Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2210.01790.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2210.01790 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:04:57.439817Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9ccb2324-bc9e-45fe-a20c-4e180e231e83 · inbound

Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs cites this paper.

Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-08T00:04:57.439817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:04:57.439817Z digest=sha256:5f2b8021e1bec86114a5262326ecd7baf0c3749b460523b683d376d6d2e30e0b

Observation bd00104a-e810-4aa8-8ada-3afa05d52c3c · inbound

Technical Risks of (Lethal) Autonomous Weapons Systems cites this paper.

Technical Risks of (Lethal) Autonomous Weapons Systems Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T19:10:56.735518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T19:10:56.735518Z digest=sha256:9d64275e3703426aab3ba7947a71e1fe375d933dbd01a35a5f2b70898f5559c1

Observation c01b1197-dea2-4fd7-8e52-628342c9f5b4 · inbound

Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies cites this paper.

Bridging Distribution Shift and AI Safety: Conceptual and Methodological Synergies Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 154

Resolution
unresolved
no resolver link, observed 2026-08-07T13:05:32.661716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:05:32.661716Z digest=sha256:93d7c88b8fea9e6687ba0992e8966c2050f0fa3a64bfab0188d74c18d039c382

Observation 5239c1e3-0a0c-4092-bdc3-736d05106210 · inbound

AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models cites this paper.

AI Risk-Management Standards Profile for General-Purpose AI (GPAI) and Foundation Models Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 1270

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:13.140632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:45:13.140632Z digest=sha256:7314989777b1c09be775aae88051aac248e41652568a27df96b2c881be63b158

Observation d89b77a1-e7f6-46bb-81c3-f37439138105 · inbound

Mitigating Goal Misgeneralization via Minimax Regret cites this paper.

Mitigating Goal Misgeneralization via Minimax Regret Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:17.268826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:17.268826Z digest=sha256:90bced134e5ca956453059b38606b2ec0f9c8436f2abd52bb6828ff7c6ae971b

Observation 3a50fbda-be39-4508-afa8-021011168d71 · inbound

Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language cites this paper.

Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T20:18:15.459547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:18:15.459547Z digest=sha256:a34463c663c0d8cc821a5f99499a2ab1dd10bf203d5f108af229bdd4bfe3d773

Observation b86a2244-c5df-4e12-bc22-fbe0aadbd572 · inbound

Semantic Convergence: Investigating Shared Representations Across Scaled LLMs cites this paper.

Semantic Convergence: Investigating Shared Representations Across Scaled LLMs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:39:46.553247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:39:46.553247Z digest=sha256:7a0b5a3ede27d91e7bf72ee460fcbf7a4f16ec623776332e000dab06f3a61741

Observation 6602d481-ab81-4b73-a57b-572c15993da5 · inbound

Interpretable Electrophysiological Features of Resting-State EEG Capture Cortical Network Dynamics in Parkinsons Disease cites this paper.

Interpretable Electrophysiological Features of Resting-State EEG Capture Cortical Network Dynamics in Parkinsons Disease Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-13T14:23:49.346727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:23:49.346727Z digest=sha256:829ca2df789c0eea28830b5e23d73f979efe30150df8d54ab078bd826dcd06d8

Observation fc413ef3-77aa-4e5c-8cb2-9ff910d1225d · inbound

Why Does Agentic Safety Fail to Generalize Across Tasks? cites this paper.

Why Does Agentic Safety Fail to Generalize Across Tasks? Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:06:00.148917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:55:38.554161Z digest=sha256:a3e7d31d7789c62177f766d9ef5d119c8d4ad9b5129e4fbe64136068b6e00dcf

Observation 56e0c30d-f7bc-4a2c-b018-96b0a27829c9 · inbound

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training cites this paper.

Spurious Correlation Learning in Preference Optimization: Mechanisms, Consequences, and Mitigation via Tie Training Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:32:24.261475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T06:30:51.812541Z digest=sha256:9031e0938b7e568b782194e864031f3697e44a8b7ce235d821acd17dc530577f

Observation d6b62d88-4d53-4aff-9442-31fa43fc5179 · inbound

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack cites this paper.

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:32:56.848195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T20:31:50.043920Z digest=sha256:07129a09cfe12974d09f4b6d52e2d059854d4d8e7b29c92e692c5350d05a2ccd

Observation 584b3b41-465c-4a39-9201-3e01e7339d36 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 224

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:54.260119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:cd263ab92e20128e086ebe94b7102bb7f84f0d9102b77ae7bcd2e8647e84605e

Observation d5631898-2d29-4a4c-a850-41288180f501 · inbound

Understanding Goal Generalisation in Sequential Reinforcement Learning cites this paper.

Understanding Goal Generalisation in Sequential Reinforcement Learning Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:50:20.839480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-25T04:49:50.034743Z digest=sha256:b577ca63c55fb7e59e8db779db4862d31ee27c9f489d597121d93944d3e37fe7

Observation 33f4e2df-1111-4edb-8047-70cf3b707736 · inbound

Voluntary Collusion with Secret Tools in Competing LLM Agents cites this paper.

Voluntary Collusion with Secret Tools in Competing LLM Agents Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:40.769914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T17:01:22.732390Z digest=sha256:e6c4d7b3e49627a5bf3f643f0a05a8011aed54bb10199cbfacc0397f47a1cd23

Observation c543ef0d-d36a-478e-9534-d9046f192b04 · inbound

Beyond Value Benchmarks: Measuring Value-Structure Alignment in Large Language Models via Symmetric Q-Sorts cites this paper.

Beyond Value Benchmarks: Measuring Value-Structure Alignment in Large Language Models via Symmetric Q-Sorts Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:09:40.879344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T12:13:12.018506Z digest=sha256:b110e9dd33bd8ea39d37d0e4d0a494394d01af10237162cc8fcd844340fe9797

Observation dd3ba819-e0b0-426e-95b0-13bf0ee1beeb · inbound

A Scalable Approach to Evaluating Moral Sensitivity in LLMs cites this paper.

A Scalable Approach to Evaluating Moral Sensitivity in LLMs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 156

Resolution
unresolved
no resolver link, observed 2026-07-12T05:44:33.099337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T05:44:33.099337Z digest=sha256:e942f9746a68eff21385e848e2992257fdf7892df8f4e51e3b21ffe72c43a837

Observation 140911e7-4a29-46ba-b69b-c6c9fecb38a0 · inbound

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary cites this paper.

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T07:20:39.148585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:20:39.148585Z digest=sha256:a6c591689d74078ecf4646c88108e8072ce49ad2df9abc870c77a6db1977c23f

Observation 1f48ad79-a499-40c0-b165-e44d061f11bd · inbound

Innocuous-Seeming Data, Latent Ideology: Ideological Generalisation in Finetuned LLMs cites this paper.

Innocuous-Seeming Data, Latent Ideology: Ideological Generalisation in Finetuned LLMs Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T00:51:18.184662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:51:18.184662Z digest=sha256:c924d72d566265b32de85d88228541613546885d35d1a37d9cc2ea248def2a69

Observation c871cfe9-711c-4206-b6f4-fbc093d770e1 · inbound

AI Value Alignment for Evolving Social Norms cites this paper.

AI Value Alignment for Evolving Social Norms Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T15:16:55.934050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:16:55.934050Z digest=sha256:397cd8418d04412dc3d49556c777f3d3afb0526d674d56f222a5900d4b663dcd

Observation 37c95f28-7fa9-4c6e-aacb-8e43d5861797 · inbound

Response drift across frontier large language models cites this paper.

Response drift across frontier large language models Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T13:58:47.006884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:58:47.006884Z digest=sha256:69bca22b53ab1fb6623486bcf4118a74261e0316f23a4fe919548c3432dd87b0

Observation 6712e349-7885-44c6-930a-99684fdaaf2e · inbound

Co-design of LLM-based preference agents: participation may drive overtrust cites this paper.

Co-design of LLM-based preference agents: participation may drive overtrust Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T06:48:46.131233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:48:46.131233Z digest=sha256:12506fe49e2f11b956a8dc246f2fd7909c58787355fa5ab67d29c648127705bc

Observation d0b3bd55-23ba-4f82-ac7f-4e85e3a62dc2 · inbound

Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives cites this paper.

Draining the Energy Commons: Self-Defeating Over-Appropriation as a Coordination Failure in Agentic LLM Collectives Goal Misgeneralization: Why Correct Specifications Aren't Enough For Correct Goals

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:12.902670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T05:36:12.902670Z digest=sha256:74e7066fcf9a28cdeb12e2e84071a6d8ccdb7db1211f6d62afd077f4655c561b