Pith. sign in

Paper Citation Record · LEDGER

The Attentional White Bear Effect in Transformer Language Models

As of 18 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2605.28639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.28639 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:43:43.792334Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch8

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4be7b14-f7d4-4dec-ac0c-ffb9c94deb40 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

The Attentional White Bear Effect in Transformer Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:53:27.035529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:6647ff4a070f7dcc05d466c3ece7669a622de1f9548db0ed62c6f7995d9629b6

Observation 0d5845af-5a7e-4218-8c30-489698a744e1 · outbound

This paper cites Discovering Latent Knowledge in Language Models Without Supervision.

The Attentional White Bear Effect in Transformer Language Models Discovering Latent Knowledge in Language Models Without Supervision

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:53:27.042291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:749ca408fdff871e9c36c485e9dac560f575c7fc08f02259c76d3e100e4a2b61

Observation b78d867d-859a-47b5-964e-45064868278a · outbound

This paper cites Https://transformer- circuits.pub/2021/framework/index.html.

The Attentional White Bear Effect in Transformer Language Models Https://transformer- circuits.pub/2021/framework/index.html

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-29T12:43:43.792334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:7c0c42de465d1008884a86dcedb94d6a16542808caec06b66ee5b021c3a61067

Observation 5451e70b-aa3c-45e6-9105-dc9960e4215c · outbound

This paper cites InProceed- ings of the 2023 Conference on Empirical Methods in Natural Language Processing, pages 12216–12235.

The Attentional White Bear Effect in Transformer Language Models InProceed- ings of the 2023 Conference on Empirical Methods in Natural Language Processing, pages 12216–12235

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T12:43:43.792334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:da0d8992fbba76d7a6545005e527696f48c96748d5d759c31ac2b78d198ae799

Observation fface771-0d51-4c28-9596-419f43f70b13 · outbound

This paper cites The Llama 3 Herd of Models.

The Attentional White Bear Effect in Transformer Language Models The Llama 3 Herd of Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:53:27.030895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:b1c2b5f22559c6bd7d0a72ebb9834f7b3daea12af0c7757c5e927288aeea5ada

Observation 1973e3d6-9c6f-46b6-8674-3a2ce0f2750b · outbound

This paper cites Mistral 7B.

The Attentional White Bear Effect in Transformer Language Models Mistral 7B

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:53:27.026047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:2c326252c7afec0c610ca887686871bfa205413569e9e421f4e5590656829e06

Observation 7548b054-74d6-4ade-9f3e-1c6268a3286a · outbound

This paper cites Redirected, Not Removed: Task-Dependent Stereotyping Reveals the Limits of LLM Alignments.

The Attentional White Bear Effect in Transformer Language Models Redirected, Not Removed: Task-Dependent Stereotyping Reveals the Limits of LLM Alignments

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:53:27.037837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:b2bc2a767b59c86f1dfe4cf440b1a61f1a0713449275e76a3d4032ac4b828b6b

Observation e2749ef2-80c9-46cd-90c1-4226907df39b · outbound

This paper cites Implicit Reasoning in Large Language Models: A Comprehensive Survey.

The Attentional White Bear Effect in Transformer Language Models Implicit Reasoning in Large Language Models: A Comprehensive Survey

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:53:27.047357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:f494557ce665378498ac5f19321ffe88c367ee9d235cbdd82943002c97a4db9c

Observation f9a9348f-6a86-469c-b49a-301f4d53c7c1 · outbound

This paper cites Large Language Model Alignment: A Survey.

The Attentional White Bear Effect in Transformer Language Models Large Language Model Alignment: A Survey

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:53:27.023930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:6ce7d96a1c099b0533ef690cb7df51d7adefa1850f18ae3aa7e808c4811a3f9d

Observation 38ef084f-1eaa-4672-a795-e38dc811c932 · outbound

This paper cites Getting aligned on representational alignment.

The Attentional White Bear Effect in Transformer Language Models Getting aligned on representational alignment

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:53:27.021462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:2831524ab49cd7f19964446debb454ae74ff5fdeb82024f1f2d848a8d108f0ac

Observation ce6b8ddf-a071-4cfd-92a3-f6e0ef9995d6 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

The Attentional White Bear Effect in Transformer Language Models Gemma: Open Models Based on Gemini Research and Technology

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:53:27.040094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:25e064cc99be20f5e39fb8e71448d45c6c24f251b7e319e95548af44aaf984a8

Observation d9fb1980-78a5-45ad-93a9-c39a5885e38d · outbound

This paper cites Steering Language Models With Activation Engineering.

The Attentional White Bear Effect in Transformer Language Models Steering Language Models With Activation Engineering

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:53:27.033239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:e5719fc5804ad46de04449c607cee174c69c1868d72813ad28693a96d5fe5561

Observation 1605d307-fbbf-4a3f-a579-9023a5d1b435 · outbound

This paper cites Erasing Concepts, Steering Generations: A Comprehensive Survey of Concept Suppression.

The Attentional White Bear Effect in Transformer Language Models Erasing Concepts, Steering Generations: A Comprehensive Survey of Concept Suppression

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:53:27.028484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:c15c89be486abeb4576a161bbb07a4e830e994005c1b8db9a69509b542f16059

Observation b05dab7c-8d4f-4e63-a7b4-5f805da9f510 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

The Attentional White Bear Effect in Transformer Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:53:27.044576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T12:43:43.792334Z digest=sha256:67f1ef3c3aee1ded3cc8a7397e23d37e50af9529125ea105b9ff8ece7488a0f9

Pith citing papers

No inbound Pith citation observations are available.