Pith. sign in

Paper Citation Record · LEDGER

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

As of 22 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2602.19101.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.19101 v2

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T21:49:10.161724Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:32:57.295159Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T09:07:47.890063Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a02c9c72-c23a-4358-9bcb-d17f97eaefa6 · outbound

This paper cites The Ethics of Advanced AI Assistants.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models The Ethics of Advanced AI Assistants

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.436079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.436079Z digest=sha256:1f26efa546918ca2c6aef3ca83ee2f20e2bfdd77205641965e19d0553b946814

Observation 476eb455-e5a0-4b8e-8197-863197103191 · outbound

This paper cites Unsolved Problems in ML Safety.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Unsolved Problems in ML Safety

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.522544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.522544Z digest=sha256:441a0ea42ade779fd2b01c3a2f7a83a9d8e4acaf8e2aa6c78c650b7865575e1f

Observation 6b90d727-52ab-4803-b6fb-de0f44b4b028 · outbound

This paper cites Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.679130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.679130Z digest=sha256:223cacc560beb221e028c8c3f189e2c22a24dabd96f5889b65c3ad07fefb1352

Observation c4bd6c0d-3735-457e-a197-5f642f9620dd · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Steering Llama 2 via Contrastive Activation Addition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.785134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.785134Z digest=sha256:d2f5ba836fbe92e7ef6d410f628bc9dbf5a47dae90aa8357d1b6f4e863bb859e

Observation bb8a2115-5feb-4a1f-994e-13e4f4bd8724 · outbound

This paper cites Qwen2.5 Technical Report.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Qwen2.5 Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.913148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.913148Z digest=sha256:8bc1dad460bbe0765491b04e3764b2aaeb9d9d50fce49ec45f4be4eb8e873cef

Observation 62bc3130-4331-42b5-8127-627d1046e46c · outbound

This paper cites Convergent Linear Representations of Emergent Misalignment.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Convergent Linear Representations of Emergent Misalignment

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.959136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.959136Z digest=sha256:ae519ab1fd38d269eaed41a5bfb07117bd3d57d687635223175a1badb9f8c47f

Observation 33404c56-3a97-45f9-a987-b6b99de887eb · outbound

This paper cites Model Organisms for Emergent Misalignment.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Model Organisms for Emergent Misalignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:10.039369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:10.039369Z digest=sha256:1ac0e8c827d54df7d4b07e88c517dce5ee3b5eb04c9067e96b1deb4ff8342771

Observation d8b0c6ed-3a59-4f74-be19-a4b844c105a5 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Representation Engineering: A Top-Down Approach to AI Transparency

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:10.093299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:10.093299Z digest=sha256:fa1441dae27c388d47714c16b437c5eae985f687b6433f36f4697b2190a36a22

Observation 6d1233c1-472e-4409-8198-6282c031fdd5 · outbound

This paper cites OLEDinstead of going to the optional work event. Neutral $$$$ I chose to watch TV on myLG 65.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models OLEDinstead of going to the optional work event. Neutral $$$$ I chose to watch TV on myLG 65

Reference 13

Resolution
malformed identifier
no resolver link, observed 2026-08-02T21:49:10.161724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:10.161724Z digest=sha256:05650a124fef354f8389ef794cade7d12b6e2393922fb7452e07c692d2c17141

Observation bef8f37c-cd31-42e8-a4b3-3b54a546f5e8 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.271327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.271327Z digest=sha256:fe4287c9ce095a44d4a306116cbf4f2ccb73c61829c10439a9fb23992c695a8e

Observation d72d5ca9-f660-4242-a164-048a0c90f11d · outbound

This paper cites doi: 10.1016/j.tics.2023.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models doi: 10.1016/j.tics.2023

Reference 2023

Resolution
verified exact
doi, observed 2026-08-02T21:54:29.151885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-02T21:49:09.350529Z digest=sha256:46491e9945c0dbed716ac042d62f7fda535dc55733eb021a621c60fd8a6a30b2

Observation ae547ce6-8ead-4da0-982f-2005f7b6f4b6 · outbound

This paper cites Refusal in Language Models Is Mediated by a Single Direction.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Refusal in Language Models Is Mediated by a Single Direction

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.092182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.092182Z digest=sha256:4912de5ba17e33d54166ff25f5921a89e09328f905846f6ed10a9eafb4cf35b9

Observation e1d58311-be77-4314-af1c-edb9ff1d7380 · outbound

This paper cites an unresolved cited work.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.167520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.167520Z digest=sha256:5eacec2183ecb6d9051f90afbb6239d08512db7086e629e854044b82a8973102

Pith citing papers

Observation b6ef70d9-560b-4770-94b7-d55c7f28f91a · inbound

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal cites this paper.

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

Reference 292

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T09:07:47.891206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T10:32:57.295159Z digest=sha256:406eacc297590dcf7457168f4b51fa91b7635929e621e78aa771edb7d6bcc9db