Pith. sign in

Paper Citation Record · LEDGER

On the Role of Attention Heads in Large Language Model Safety

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2410.13708.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13708 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:11:18.246898Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T17:35:51.334325Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ed3a0af8-9b98-495b-a1ef-b264678b1046 · inbound

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems cites this paper.

To trust or not to trust: Attention-based Trust Management for LLM Multi-Agent Systems On the Role of Attention Heads in Large Language Model Safety

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:32:17.349169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T11:30:47.877793Z digest=sha256:a675e9dd7bf6398a6584bf45702915b7417b57e10a4ef71e3a70ff436a3b93f4

Observation 3ab5dd33-7b72-478c-8828-a72b8a526888 · inbound

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents cites this paper.

SafeMobile: Chain-level Jailbreak Detection and Automated Evaluation for Multimodal Mobile Agents On the Role of Attention Heads in Large Language Model Safety

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:11:18.246898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:11:18.246898Z digest=sha256:098c8c115c3928a902e635a1845b405c071d6ddb9681a6ee863fa1760a728cc3

Observation 45bcdb12-ec21-4f9e-93f2-3216288e7406 · inbound

Boosting Parameter Efficiency in LLM-Based Recommendation through Sophisticated Pruning cites this paper.

Boosting Parameter Efficiency in LLM-Based Recommendation through Sophisticated Pruning On the Role of Attention Heads in Large Language Model Safety

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:54:18.141878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:54:18.141878Z digest=sha256:1db019545db406836a9655570bb78f23fe51b0bcb88e2b45aa5373a98b191830

Observation 838f11fa-4a06-497f-abe6-414a313cc94e · inbound

Soft Head Selection for Injecting ICL-Derived Task Embeddings cites this paper.

Soft Head Selection for Injecting ICL-Derived Task Embeddings On the Role of Attention Heads in Large Language Model Safety

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:41:59.765500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T02:41:52.906516Z digest=sha256:9abe09bc1f14b4f8166c9ccb593a0b002ced452cc7f52dbcd02a1968cbbb7189

Observation 3f6b9687-55ac-49c7-8e14-5ecae468d82c · inbound

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering cites this paper.

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering On the Role of Attention Heads in Large Language Model Safety

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-04T11:19:51.206012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:19:51.206012Z digest=sha256:c7486efd80aae1c340096ba6b2a0f1f6b887daf9656fca221140f0c62b6b7064

Observation 84a6d79a-0654-4057-9f86-0d522e155444 · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems On the Role of Attention Heads in Large Language Model Safety

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:16:03.952450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:5ef209dd4dc174fbaa2899b70c85f883e53937114a7ec954e50875217368acf0

Observation db8833b2-98aa-4b3b-a9f4-1031a28c3e12 · inbound

Why Do Large Language Models Generate Harmful Content? cites this paper.

Why Do Large Language Models Generate Harmful Content? On the Role of Attention Heads in Large Language Model Safety

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:21:04.490884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:31:13.545599Z digest=sha256:f74147fb0576f4476677ebee20c38ddb18319337227ad0980790281e43da87db

Observation e78737db-90cd-47af-9522-f14ca3af9047 · inbound

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs cites this paper.

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs On the Role of Attention Heads in Large Language Model Safety

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:01:27.533964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T08:34:14.310656Z digest=sha256:6edcbb07948f9cbe9be05340552332e544fee98936a81cc3235f1dfebd05d21b

Observation cc737df7-85e7-4416-ab3a-9c04ae0c2912 · inbound

Large Vision-Language Models Get Lost in Attention cites this paper.

Large Vision-Language Models Get Lost in Attention On the Role of Attention Heads in Large Language Model Safety

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:26:10.092015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T11:54:01.224588Z digest=sha256:ed69ba99bc1e68a5f86133131e96c64e3fcfcfa100ab7f3b6ac0c64f476848c3

Observation 30323868-1d1e-4566-898e-a2ae814bef17 · inbound

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models cites this paper.

Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language Models On the Role of Attention Heads in Large Language Model Safety

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.457378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T13:25:40.986336Z digest=sha256:26d017fef02d6f60781d27e5b50868e72b8fe9b787dde328d311069d15f945e3

Observation 67aa9938-d5ba-4348-b004-945de94a1d1a · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models On the Role of Attention Heads in Large Language Model Safety

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:35:51.335775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:134d93813b24f14f101485f1b0e874ae1bb7dc20105d3f4dc899abde0d68236b

Observation 4edb1c6a-47b3-479b-8f9d-a0718c5ca62c · inbound

How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Program Repair cites this paper.

How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Program Repair On the Role of Attention Heads in Large Language Model Safety

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T01:17:54.450344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:17:54.450344Z digest=sha256:f340e30bdaabe79145a115bc54d6ffa7b0ead4ba31242a957a8b1777e5374816