Pith. sign in

Paper Citation Record · LEDGER

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

As of 9 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 6 inbound Pith citation observations for arXiv:2507.03662.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.03662 v1

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:12:32.717267Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T07:46:16.678202Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:47:29.934082Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 62f4367f-cd43-4297-967f-69d6ae87b075 · outbound

This paper cites URL: " 'urlintro :=.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:30.497834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:30.497834Z digest=sha256:2a6e79922198825e015c2b6f4ee7d21ee664de9cf4fcc019890e492f43fef771

Observation 25dbe0fd-3487-41a7-b685-04a8f11c4335 · outbound

This paper cites write newline.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:30.565244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:30.565244Z digest=sha256:103bfdb9180bb07a7d2899d10c7e7a291322b8072fbdaa84502f44f3ea256682

Observation a5c470c2-9592-4749-a3eb-23587dfd7d9d · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.140998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:30.799717Z digest=sha256:a8d96b28cd932ffaefef686c710378cd2428f038e10b90f8fbb5880af8a95537

Observation ca7bb3f9-e6e4-4ac1-8ab9-ecfc7afcde91 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Constitutional AI: Harmlessness from AI Feedback

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:30.963700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:30.963700Z digest=sha256:4251c92024ec860fe4cfe089234b5a1e95a911579399c217746c100d88b60d75

Observation 97a054a0-bc05-4b2c-a4df-343e9ba49a7e · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.133118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:31.086504Z digest=sha256:d7fa3152873f01b26b349d43651a3f4f13cb6d66267219c48016414df75632ba

Observation b299773b-7daa-495b-bb1a-9afbb6bba5c0 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:31.194346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:31.194346Z digest=sha256:a9ae8acb87af81f789cee0714260b705262a8707320e577e7cc65a107a142d28

Observation 625618f7-5876-4f5c-9d46-a9ca13d9fed9 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.125055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:31.342729Z digest=sha256:382aa0ed320b3cdb6c8a9b037629fc484eb2e339732d3e9aa3ef1c27f8c40c60

Observation 3cc5fc6c-864d-44a4-b643-d10f4d2acc4a · outbound

This paper cites Alignment faking in large language models.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Alignment faking in large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:31.417382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:31.417382Z digest=sha256:99e1d057d8e0cc1cac28784b9d8b6b996bb77c733b5f93c4aa82dc1725d45c56

Observation cd9b383d-6907-4e68-a577-33a83da7a7c5 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Qwen2.5-Coder Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:31.510017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:31.510017Z digest=sha256:dca032ae2ff291980b00ce663ffce7d82b14b2a594017916e56e7570a6adb3d6

Observation 6ef3bc1d-de03-4454-8877-edee4a792cf4 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.116840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:31.619729Z digest=sha256:e198222ca3e50ea1e7eacd9659040150cdaf436d66dc919bb203b0facd91da74

Observation e5f17dd1-7cd4-439e-a439-a55953373809 · outbound

This paper cites What Happened in LLMs Layers when Trained for Fast vs. Slow Thinking: A Gradient Perspective.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs What Happened in LLMs Layers when Trained for Fast vs. Slow Thinking: A Gradient Perspective

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:31.713901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:31.713901Z digest=sha256:3febf31b0ccd630e5c976ee5f8ce7a040b43e3ff473c52f456d811e85ca1874d

Observation c72ec5de-4211-420b-96be-6e2b12e39de5 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:31.818876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:31.818876Z digest=sha256:2169c0a762ec8d24e3f2b6da740257acc785a5663cf24c94bedcfab8c70bbb9a

Observation 4352885e-eeab-4b84-bd8f-08c00e0f0c34 · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:31.934646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:31.934646Z digest=sha256:096f8675da7cf73398e15002c401d795b8f5f709fb1c3e97762d4da30d87e350

Observation 9295086e-065f-4692-bae0-5645d92b2092 · outbound

This paper cites The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.082534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.082534Z digest=sha256:51553e206f8dfbbefb71577d092caa954a749f2f833d5d527bb909b41e1b64fe

Observation 89996c97-33f7-4b19-af30-445a6f932ffb · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Frontier Models are Capable of In-context Scheming

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.170069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.170069Z digest=sha256:59812e7c9604823f2b882af054e7b3f16a29b1de3ad6ea88774da732d8d3a10c

Observation 7b2a8d3a-8b9f-4bc4-a50d-dc6304d93ee6 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.103446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:32.292132Z digest=sha256:800ffa6f540e1f6b5332a9e371019a2be6300319c2610a72688e8443a3338041

Observation b416694e-7921-4807-93b9-12890c8c63dc · outbound

This paper cites The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.398404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.398404Z digest=sha256:91eba43c4eff2a21d44ca4a8fcd5245bdf91104c6ae358d85f636d871494943b

Observation 939f9b4c-eefc-4e53-a938-b2c1c323f77b · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.095322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:32.567511Z digest=sha256:5883fc56579265fdafa1e3ef2aa2d7fab1fac1a88813abe1de70d7d755c8dfaa

Observation bafb2894-8f73-459c-8955-986e3df6e39a · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.643003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.643003Z digest=sha256:456f347cfc5443768d3b21d9aab945b6858e0a13a6f78cfcacec86ec627f95c1

Observation 65189b11-ca14-4d88-a6e2-8eccd78c7420 · outbound

This paper cites Large Language Model Alignment: A Survey.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Large Language Model Alignment: A Survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.681144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.681144Z digest=sha256:ba4b67f59d2914a66866dc2df642b8a47ab4923e3a6f580094afcbb43f874c0d

Observation 9f9a9da4-00a7-4c63-99f1-b22f0e2864bb · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.080992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:32.699969Z digest=sha256:7cf3e44ab2206bb02f748305c5bc4fc3def225ec44e4c469ec77c53cc1881dea

Observation 977ae365-45ac-4b9a-a33d-300974733ea9 · outbound

This paper cites Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:12:33.072891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:32.703259Z digest=sha256:5817901655cbcd9ebac6e48bf63c61f59dd94410747374a57d66494a5cc31551

Observation 197cdb0f-11bf-4a8f-b2f3-cdaac7a31d69 · outbound

This paper cites Tradeoffs Between Alignment and Helpfulness in Language Models with Steering Methods.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Tradeoffs Between Alignment and Helpfulness in Language Models with Steering Methods

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.706277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.706277Z digest=sha256:4fd38b7d66688b397722b85ca702c837363ad1008356c7e89b3bace9b2095b07

Observation bdee2e24-4579-4b38-a4bc-fb759751952f · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.709279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.709279Z digest=sha256:8c3d298e44f5f22ecd4ac5a59bd1ba4a9b1d4e10a3b7824dfc326019b8d8911f

Observation 21face3e-8415-4880-af1a-fe573448aef9 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.711845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.711845Z digest=sha256:4cfd68ce125d1750705bee6d2bf8109189309871087fb3c8b7c1f8b5c31cbaed

Observation 11ef7507-24b9-4b9b-8132-d4782471ece7 · outbound

This paper cites an unresolved cited work.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:12:33.063688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-06T20:12:32.714663Z digest=sha256:f3c72e29b8786e36b99bca107d8c66b02498468a682d506768f705af4bf58776

Observation 3798a0a1-33f5-473c-9ba6-a532c6038529 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs Representation Engineering: A Top-Down Approach to AI Transparency

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:12:32.717267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:12:32.717267Z digest=sha256:f4dc7bc83f4b035894a04ba15d8ccaf08b92df061cc201b66bc6f298b015f254

Pith citing papers

Observation bb166d1b-0ba0-4538-9556-df44c085b1e4 · inbound

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer cites this paper.

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:39:26.793422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T20:36:44.104197Z digest=sha256:927ac355969f5d6f037cfde474e88730af182e9dd022abbbbe30b00e72a75688

Observation a9e7daba-9750-43ab-aaa4-3a5d18d69f6c · inbound

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence cites this paper.

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T20:22:37.266421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T20:13:53.972585Z digest=sha256:82bf7139efa77504097b99c2171c0eca01b06502ae24ba9c108fed4c188c23bb

Observation 54b7d6bf-2f22-4b85-a8b1-d71a6906e547 · inbound

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective cites this paper.

When Behavioral Safety Evaluation Fails: A Representation-Level Perspective Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.094369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T20:04:17.744876Z digest=sha256:669fee60d2bb2ce9c4d7c4b7864d69a631a36f770ffb10adb8da4dea2ebb3a7e

Observation 9f2ba424-287a-47fe-808c-60bbae68247a · inbound

Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating cites this paper.

Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T00:47:29.935594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T17:03:33.199645Z digest=sha256:0be4bd4ddc3184a2b55dbee8fa76e134076c71294fbd6913d7283e98764e9687

Observation dae3c0ef-8390-43fa-9503-c5e75a1078e8 · inbound

Transplanting, inverting, and preventing a misalignment persona: method-conditional emergent misalignment in Qwen2.5 cites this paper.

Transplanting, inverting, and preventing a misalignment persona: method-conditional emergent misalignment in Qwen2.5 Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T18:20:40.287006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T18:20:40.287006Z digest=sha256:743a24450922eb4268519fa84e7113d2a7d1bf0889abe2a80657acc06d2665bb

Observation 71991987-2695-4185-8692-2fddc292ff96 · inbound

Emergent Misalignment Recruits a Pre-existing Persona Subspace cites this paper.

Emergent Misalignment Recruits a Pre-existing Persona Subspace Re-Emergent Misalignment: How Narrow Fine-Tuning Erodes Safety Alignment in LLMs

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-01T07:46:16.678202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:46:16.678202Z digest=sha256:f0a70b2ce78687ca4fd8e93171c411a95f31fa680f4b41300509ed426950ddf2