Pith. sign in

Paper Citation Record · LEDGER

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2602.19101.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.19101 v2

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T21:49:10.161724Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:32:57.295159Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T09:07:47.890063Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a02c9c72-c23a-4358-9bcb-d17f97eaefa6 · outbound

This paper cites The Ethics of Advanced AI Assistants.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models The Ethics of Advanced AI Assistants

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.436079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.436079Z digest=sha256:28e52781f5901cdc57229f5b2065110e02fbc7eb1d5abeb79b3bdc22328f06d8

Observation 476eb455-e5a0-4b8e-8197-863197103191 · outbound

This paper cites Unsolved Problems in ML Safety.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Unsolved Problems in ML Safety

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.522544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.522544Z digest=sha256:a8309f9112a0c278bb192c380dc224d3cad9e9263d05ad55d2e3d3d5bca43290

Observation 6b90d727-52ab-4803-b6fb-de0f44b4b028 · outbound

This paper cites Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.679130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.679130Z digest=sha256:a05caa3afdc8f37f8e6469623076da49b99e0d7d6c273ea5dabc4d5353a20fc4

Observation c4bd6c0d-3735-457e-a197-5f642f9620dd · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Steering Llama 2 via Contrastive Activation Addition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.785134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.785134Z digest=sha256:87e8ad2a265b6c2f66047c742e612cd8d8617208ba7bb52a75022241cc34a49b

Observation bb8a2115-5feb-4a1f-994e-13e4f4bd8724 · outbound

This paper cites Qwen2.5 Technical Report.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Qwen2.5 Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.913148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.913148Z digest=sha256:4dd879373a59d1c18bf71d2a47ece7689f31777dda66faccffa9ec173a0ae83d

Observation 62bc3130-4331-42b5-8127-627d1046e46c · outbound

This paper cites Convergent Linear Representations of Emergent Misalignment.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Convergent Linear Representations of Emergent Misalignment

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.959136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.959136Z digest=sha256:df2e51a190bad52deb867d43547621f68d45b39393da76a6701cfec15216f54b

Observation 33404c56-3a97-45f9-a987-b6b99de887eb · outbound

This paper cites Model Organisms for Emergent Misalignment.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Model Organisms for Emergent Misalignment

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:10.039369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:10.039369Z digest=sha256:80d781caffe2798e4f1e3be45844bdb968a04fcc588c4079ce39229acb873db0

Observation d8b0c6ed-3a59-4f74-be19-a4b844c105a5 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Representation Engineering: A Top-Down Approach to AI Transparency

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:10.093299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:10.093299Z digest=sha256:f64be77bef4c1323b80cea81560a251d63dfe4b1dd447936e0b9f0d642e9fad6

Observation 6d1233c1-472e-4409-8198-6282c031fdd5 · outbound

This paper cites OLEDinstead of going to the optional work event. Neutral $$$$ I chose to watch TV on myLG 65.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models OLEDinstead of going to the optional work event. Neutral $$$$ I chose to watch TV on myLG 65

Reference 13

Resolution
malformed identifier
no resolver link, observed 2026-08-02T21:49:10.161724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:10.161724Z digest=sha256:83e5099d0644154b20f587909a4816f847b81ea176653e4f4f30d5340c6230ed

Observation bef8f37c-cd31-42e8-a4b3-3b54a546f5e8 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.271327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.271327Z digest=sha256:43c62156ffbdb20ecfe2b6ebf318e270f54fa41b856d1b64085fe116eef2730e

Observation d72d5ca9-f660-4242-a164-048a0c90f11d · outbound

This paper cites doi: 10.1016/j.tics.2023.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models doi: 10.1016/j.tics.2023

Reference 2023

Resolution
verified exact
doi, observed 2026-08-02T21:54:29.151885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-02T21:49:09.350529Z digest=sha256:2e44e192412717194c939448a0797b95af08bf43442e41b644de4b175bb0b809

Observation ae547ce6-8ead-4da0-982f-2005f7b6f4b6 · outbound

This paper cites Refusal in Language Models Is Mediated by a Single Direction.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Refusal in Language Models Is Mediated by a Single Direction

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.092182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.092182Z digest=sha256:966c5c687121559528e9b4b67123dccfe68db6633343efca08aa69ee92605e87

Observation e1d58311-be77-4314-af1c-edb9ff1d7380 · outbound

This paper cites an unresolved cited work.

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T21:49:09.167520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:49:09.167520Z digest=sha256:eef19f7726e4b75c2d64c18f249ab1974e295a2e83f114128048d2eecc8bb75d

Pith citing papers

Observation b6ef70d9-560b-4770-94b7-d55c7f28f91a · inbound

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal cites this paper.

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

Reference 292

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T09:07:47.891206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T10:32:57.295159Z digest=sha256:fc0067a2b58f92f057c94218f6bf70992fdd66016c0ed0eceafb7878bd1a1bc8