Pith. sign in

Paper Citation Record · LEDGER

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

As of 9 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 2 inbound Pith citation observations for arXiv:2607.07903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.07903 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T15:42:46.392593Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:59:10.702144Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-08T16:59:12.165147Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact12
  • verified fuzzy48
  • unresolved2
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 680e0b50-e187-4317-a51e-20bdf9bb804d · outbound

This paper cites Language models are few-shot learners.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Language models are few-shot learners

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.566944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:f6ac0dc516c888ba333928a9455dd14f9c08b98c9ff815ea5bce5418a1bf6073

Observation 8383f39d-3cce-4125-ab39-1861964c3d2c · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Attention is all you need.Advances in neural information processing systems, 30

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.568568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:7c6f1c4e64942cfc12fdbf4b6c76dd2e2b2d122d05408b4e98a9876f45edec0c

Observation 9c69b763-8d3c-48a5-b556-67661ba8f832 · outbound

This paper cites Explaining and Harnessing Adversarial Examples.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explaining and Harnessing Adversarial Examples

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.187453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:41d721e957dacba75196e06810334655a6bb49dffec15db15b5a133359e33b8b

Observation f1c3ed2e-5048-4696-a2b7-60fac5d7d455 · outbound

This paper cites Towards deep learning models resistant to adversarial attacks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Towards deep learning models resistant to adversarial attacks

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.563775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:05d05ea398e8eed8a6f7e846da34443ed94b31f228cf0ab691af7f0311014b00

Observation 59b98507-bd75-4512-8aa5-5b54239ac5ff · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.185136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:01c6c908ce7a831943568ccba6d94cb754336402442dede7d3322f324ee0dc24

Observation c27c8f31-64b1-46b0-9b8a-559bb29ca91a · outbound

This paper cites Training language models to follow instructions with human feedback.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Training language models to follow instructions with human feedback

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.570237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:cebf409610eceee7c5bb69938c3a92b6b60df1a0593de8d613ac5e54c41ade58

Observation 1b58cf5e-ea7b-4e69-8c94-62bf1c701a03 · outbound

This paper cites Bridging Interpretability and Robustness Using LIME-Guided Model Refinement.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Bridging Interpretability and Robustness Using LIME-Guided Model Refinement

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.205521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:37007b1c0bc04dbfa7ee2bdf2c45f0a1e834596fbf740e952ab6218ccd7dc99c

Observation 169b105d-5578-4346-8fd4-1a8ad474fdfd · outbound

This paper cites Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.571948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:466c1aa943657517d3e07349ab63e042a40893c759b7f5ac5999224bc93f78d3

Observation 37d3f899-fd79-4665-9b35-54a34dd9b37f · outbound

This paper cites Explainability-guided defense: Attribution-aware model refinement against adversarial data attacks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explainability-guided defense: Attribution-aware model refinement against adversarial data attacks

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.607238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:a0ae7901faa72a0ea9fda362a2cbc84ad6f213d82caa10b3cda4ef4a698bf94f

Observation 601bde07-b71b-497b-ae28-47824b59c8f6 · outbound

This paper cites Representation learning and nature encoded fusion for heterogeneous sensor networks.IEEE Access, 7:39227–39235, 2019.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Representation learning and nature encoded fusion for heterogeneous sensor networks.IEEE Access, 7:39227–39235, 2019

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.622563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:200aa9b6705f5a3c3b319588e90012c9f60b623893872465934d752e7cdc90e2

Observation 08ace148-a560-4c8f-befa-881010b32548 · outbound

This paper cites Congestion aware dynamic user association in heteroge- neous cellular network: A stochastic decision approach.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Congestion aware dynamic user association in heteroge- neous cellular network: A stochastic decision approach

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.648376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:58925fdf3894056592a6cda4bf0f2527040e42bd9c21a12413265c55d43a8ef7

Observation 38a14a71-814c-4758-9711-2cb9b5ec9305 · outbound

This paper cites Explaining the behavior of neuron activations in deep neural networks.Ad Hoc Networks, 111:102346, 2021.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explaining the behavior of neuron activations in deep neural networks.Ad Hoc Networks, 111:102346, 2021

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.581786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:a9e850b809fa50096adac92f68afa8c631fc3848d143bf36065cac586409ae72

Observation 2e3ff8f1-6604-4e63-820e-64a19a80671b · outbound

This paper cites Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.596843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:0ee949f17c7b9b82a6447d0f595d613dcaba4fd312085437c72754a4241e98e4

Observation 2a0bb3cb-2c49-49a7-b1a4-cd8b05c3e35b · outbound

This paper cites Deep reinforcement learning based computation offloading for mobility-aware edge computing.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Deep reinforcement learning based computation offloading for mobility-aware edge computing

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.573528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d3e4afe56af8317def6ef7a0ea85aede101f65c5a10158cae26108e1a5541c15

Observation 01fb565e-efb7-4e11-966d-9208c0af634f · outbound

This paper cites Improving robustness of deep neural networks via large-difference transformation.Neurocomputing, 450:411–419, 2021.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Improving robustness of deep neural networks via large-difference transformation.Neurocomputing, 450:411–419, 2021

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.577671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e04426164c146d67fb1ff35182cc91966cbf1d00bc2dcef87f321ba8088270ba

Observation be1049dd-cf8c-461f-b0b2-d8dcf7021290 · outbound

This paper cites Looking beyond content: Modeling and detection of fake news from a social context perspective.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Looking beyond content: Modeling and detection of fake news from a social context perspective

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.646719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:66ac23dcee2e736538b5e6adc0fd58b2758fe96c0fd927bcf6237446c54c9c3d

Observation 409e11a4-216b-4dfc-845d-1bb5de8c45be · outbound

This paper cites Layer-wise entropy analysis and visualization of neurons activation.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Layer-wise entropy analysis and visualization of neurons activation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.579372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:823f15c78b0a80a5c6d30c68a946ea57e8810d3a2dfe4b34ea734958aa8116b8

Observation 65310eba-0524-4ea7-98df-0a2f9476d6ff · outbound

This paper cites Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.189844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:7fb7ff64f8c3b8f84e0cc0f92565fb03df0d9ddafdf6070a92b29c293d1e1a29

Observation 47b283e7-d437-4037-ad18-9d223467b674 · outbound

This paper cites Explainability- driven defense: grad-cam-guided model refinement against adversarial threats.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Explainability- driven defense: grad-cam-guided model refinement against adversarial threats

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.650203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:c057475f7a592e8ed4f4ba50cfa77e000e1fd14b1f8bf311cc224b00b701844b

Observation b2927e6b-2bb0-4846-8c83-fea33baa0375 · outbound

This paper cites Expert-guided explainable few-shot learning for medical image diagnosis.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Expert-guided explainable few-shot learning for medical image diagnosis

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.565360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:caf22ad9ac0c3f599ef682c08ca3be3cfb5229815cb8ca61e8c181e239d0a4f6

Observation 44002a84-b16f-4c4b-b3a4-c475c6d416dc · outbound

This paper cites GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.196757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b67a29c971e2ca39a292db099afff7c66e4b61a5f7692249264cfe6b040cedf3

Observation 7d378a65-33d7-4f85-832f-c5ff4b39d6ba · outbound

This paper cites Toward carbon-neutral human ai: Rethinking data, computation, and learning paradigms for sustainable intelligence.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Toward carbon-neutral human ai: Rethinking data, computation, and learning paradigms for sustainable intelligence

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.575486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fbaa29bbd6f75ffc0ccedf92c4bb5be6f2c8bbe6fb8e4442fad40c6f7d25810f

Observation 8a5350e3-9bc5-4180-a294-53ad400d5a1f · outbound

This paper cites Expert-guided explainable few-shot learning with active sample selection for medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2026.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Expert-guided explainable few-shot learning with active sample selection for medical image analysis.IEEE Journal of Biomedical and Health Informatics, 2026

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.643323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:4fba937fc5782dec04963a838276a60d8d213215df7f5704ce0afeb4a931f078

Observation cd9bbc7a-77e1-4460-9b12-369abd128e75 · outbound

This paper cites Acting flatterers via llms sycophancy: Combating clickbait with llms opposing-stance reasoning.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Acting flatterers via llms sycophancy: Combating clickbait with llms opposing-stance reasoning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.645010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:f92ce518388d2f090fa54db83de021b3c4ce38f1529dc11ddf6c9e42afcb25bf

Observation 56c2b89e-d054-4708-9420-f080785bbee7 · outbound

This paper cites Bridging symmetry and robustness: On the role of equivariance in enhancing adversarial robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Bridging symmetry and robustness: On the role of equivariance in enhancing adversarial robustness

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.651911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:4d4635904643144bd3bbaf170a5ed1899a9b8d766f5b3e39814bc6d6310bf504

Observation 1ef30688-2eb1-4f7a-9580-3f2f42d3d60f · outbound

This paper cites Channel- selected stratified nested cross-validation for clinically relevant eeg-based parkinson’s disease detection.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Channel- selected stratified nested cross-validation for clinically relevant eeg-based parkinson’s disease detection

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.638178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:f944f9b0245b32f083785c55286933582aa5fc64edd4c9e3462a38df6f31a724

Observation 21ce97ec-117e-4049-b15b-71a8d95945cc · outbound

This paper cites Winsor-cam: Human-tunable visual explanations from deep networks via layer-wise winsorization.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Winsor-cam: Human-tunable visual explanations from deep networks via layer-wise winsorization.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.636392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fa64c6e04499a332b4cac794a887475d34dd7c54f38be29e5818a90c3dd4e37a

Observation 4d088faa-8f00-4701-a457-ed8ef819dd97 · outbound

This paper cites Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.639983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d158347cd7268037e101de682b68fc610da34cc741ca13540a66e3109db9e22c

Observation 9826756c-7d04-4d67-abcb-0816e0db9894 · outbound

This paper cites CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.210188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e91f3441112c4c35f9df99552b80fbc43062823b6b12c00b4f782aae780d31f6

Observation 125e7890-7f5b-484f-9f52-1cd7f5b3d5e7 · outbound

This paper cites Zoom in: An introduction to circuits.Distill, 5(3):e00024–001.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Zoom in: An introduction to circuits.Distill, 5(3):e00024–001

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.632952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:bb764f687b9d11c884cb1c207f27bfde0e18e05045c7684261b3c27b551e1a16

Observation 0e02460d-b20c-4e45-877c-28c239cd849f · outbound

This paper cites A mathematical framework for transformer circuits.Transformer Circuits Thread.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs A mathematical framework for transformer circuits.Transformer Circuits Thread

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.634679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:a91c97d28f45649d1ae086ff030b660f2777d931dfb540e57344c89464bd3715

Observation 5d3ef84b-4c08-426c-bd4a-f016d587c3ad · outbound

This paper cites an unresolved cited work.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-07-10T15:47:23.641528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:cc801fcdb946f7d15677f3e1b6b342f5239dc847b45e60e7d0ba57ce9bbdc14a

Observation f64d98e3-b05d-4a52-ac51-21ffc217c163 · outbound

This paper cites Axiomatic attribution for deep networks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Axiomatic attribution for deep networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.627734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:a459fa34cf1ff9e322d66505fba8b19a991ebe127925ba3de78ee85e772dbb4c

Observation 19c3e22f-0b0f-4cb3-bd67-a265097642f9 · outbound

This paper cites Intriguing properties of neural networks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Intriguing properties of neural networks

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.207919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:272db17714fbcdb20f798ec438daf4590752c2355b31548ef34a0a7eb53f99dd

Observation 9d1f8654-46f5-4d79-b9f4-af89aeedf039 · outbound

This paper cites Towards evaluating the robustness of neural networks.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Towards evaluating the robustness of neural networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.629384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:8878e30706c4b9c899b79f07812ee0bdc23ca535483189be4a08765c19998a66

Observation 696d66d3-169e-44dc-b7eb-64a09802ec10 · outbound

This paper cites Visual adversarial examples jailbreak aligned large language models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Visual adversarial examples jailbreak aligned large language models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.625900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:589ee7ef601dd586ee30b6e894eb1d0999596ae454337811afc820af99ec5518

Observation 418d9386-7a86-4bdb-ac2a-3bacb33122ee · outbound

This paper cites AutoDAN: Interpretable gradient-based adversarial attacks on large language models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs AutoDAN: Interpretable gradient-based adversarial attacks on large language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.631154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:13dcee60af05dd20cea45aaf92f0dcce6da363ce8abd7ee72b6e84132ce59a79

Observation 2f013e77-14d3-4bad-8f75-077096bd4231 · outbound

This paper cites DOI:10.23915/distill.00010.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs DOI:10.23915/distill.00010

Reference 38

Resolution
metadata mismatch
doi, observed 2026-07-10T15:47:22.952574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fe5d94fab5c48c8b75a0b1d5ac733f7fef489cc6784a58915d728536a93c87be

Observation 834c07d4-92ff-4961-8321-d67b623b6f49 · outbound

This paper cites In-context learning and induction heads.Transformer Circuits Thread.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs In-context learning and induction heads.Transformer Circuits Thread

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.653540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:511ac278e806c0b626c31df56ea3f464dee62c5abed5f728734475102444dead

Observation e9fd6d66-7bb3-4da6-90b3-89fb35b85ab4 · outbound

This paper cites Sparse autoencoders find highly interpretable features in language models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Sparse autoencoders find highly interpretable features in language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.619321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e769709581b3ed97982558c17fc536c426da169031e1d1bcb835184dc2860ce2

Observation 009f8fb8-8db9-4acb-8ff1-5628c34fbb89 · outbound

This paper cites Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.203361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:0e5eefe89693c3f495da0e2412340f234269100e5496f163ff778c26a19ef51d

Observation 5905a272-3c72-45d8-b89e-9b09ca947f98 · outbound

This paper cites A unified approach to interpreting model predictions.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs A unified approach to interpreting model predictions

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.620873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:c894fa7e855e1eddb522460e544ef6a678b770752f7f43245c4857fa140631e3

Observation b3b1e879-d152-4107-99f6-b041cf7d6430 · outbound

This paper cites why should i trust you?.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs why should i trust you?

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.624246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:8ec7ae96bc962e33cc8f72d02a842935aa342b2cfffeb059981122d4313bd6bb

Observation c1ac42e5-49e7-4172-b397-bfc69a7a492f · outbound

This paper cites Attribution patching: Activation patching at industrial scale.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Attribution patching: Activation patching at industrial scale

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.614430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:4fc97d7e841ca6529d01e535805139db8749808a59d391c007d5af36231585fa

Observation 9e490a52-f32a-4541-ab21-a0c902bcf5e2 · outbound

This paper cites Causal abstraction: A theoretical foundation for mechanistic interpretability.Journal of Machine Learning Research, 26(83):1–64, 2025.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Causal abstraction: A theoretical foundation for mechanistic interpretability.Journal of Machine Learning Research, 26(83):1–64, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.616097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:f9d5a629e96251ee6c0f34aad584f1ea2ed87775ea6a082758f75c65e798fc32

Observation 50fd132c-5ad3-4a3e-9f9f-c77909e573fc · outbound

This paper cites Investigating gender bias in language models using causal mediation analysis.Advances in neural information processing systems, 33:12388–12401.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Investigating gender bias in language models using causal mediation analysis.Advances in neural information processing systems, 33:12388–12401

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.612655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:62a175e59eddd075ee280492a3e1780aac24f8609eecb5879e7a4df8ac456185

Observation ce825ff0-56c2-4612-98cd-990467af4351 · outbound

This paper cites Locating and editing factual associations in gpt.Advances in neural information processing systems, 35:17359–17372.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Locating and editing factual associations in gpt.Advances in neural information processing systems, 35:17359–17372

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.617762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:682ae15f1196eb973e4eb0b4f9ef5df90b99b42f9061e82a50cccc1f8ebbda9b

Observation 3185777b-bdb9-46c8-8c98-75c128dab86c · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs LLaMA: Open and Efficient Foundation Language Models

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.201157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T16:08:17.350515+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:1a35a6791220597ae677a516ba33469093fef9f8776c5b52f0be9897e348fc52

Observation e8b718b0-36eb-476e-9798-19ee400e2889 · outbound

This paper cites Jailbroken: How does llm safety training fail?Advances in neural information processing systems, 36:80079–80110, 2023.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Jailbroken: How does llm safety training fail?Advances in neural information processing systems, 36:80079–80110, 2023

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.605525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e79a57614b0e4793f2ad62364f07082d7495250e08c8e2628f9c6f5b78decacd

Observation 167d61a9-581f-43ff-a259-c3e74f7c7326 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.198943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:68e94595ba6c157ac75adf2d1dbc34b8c3a83126a3436e61226909dba735b75b

Observation da099723-6c95-48f6-8c9a-fb37b679a2fb · outbound

This paper cites Adversarial examples are not bugs, they are features.Advances in neural information processing systems, 32, 2019.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Adversarial examples are not bugs, they are features.Advances in neural information processing systems, 32, 2019

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.609025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:1e1e904b0f8328df48dc6c88cd2a9d5115e1c63f7a604ac33398edfc43950b92

Observation f6dd829d-a2b0-4d9b-8998-b4a26c6e62d7 · outbound

This paper cites Learning important features through propagating activation differences.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Learning important features through propagating activation differences

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.600092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:e58fa362d18d2dfa25052b5c3c26865fcad5d7912f091a07aa3790febf68e089

Observation 979d68bc-b92e-48b9-8b2d-8381f37b570d · outbound

This paper cites Many-shot jailbreaking.Advances in Neural Information Processing Systems, 37:129696–129742.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Many-shot jailbreaking.Advances in Neural Information Processing Systems, 37:129696–129742

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.601830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d5cb6b4d4e5ca542a52432243dc01767747413ec12457d29ba0367aec695bdf1

Observation b3167a1e-fa7b-4ee4-8bd8-ab8dabe259fe · outbound

This paper cites Cambridge university press.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Cambridge university press

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.603584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:4df9260dc86b7711caa8be1dcd5abd3500d7e8d3f89335418151235c2abb7cac

Observation c9046830-c70d-4c03-9f5b-7f95a9a6d6a1 · outbound

This paper cites Towards automated circuit discovery for mechanistic interpretability.Advances in Neural Information Processing Systems, 36:16318–16352.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Towards automated circuit discovery for mechanistic interpretability.Advances in Neural Information Processing Systems, 36:16318–16352

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.610848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ad0d26abdcfc0cd39ee1daa0b102e0c70036a388b64c63126ce14cdcc1cac4a7

Observation e795c348-4efc-4d42-8f9d-5af2832ea141 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.192014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:6c1f0a96947e1e40c035454e051ce8377240845f4d43f9d6b93c40afab16143c

Observation 7f1d3f38-1633-4655-bdb2-ef449938f344 · outbound

This paper cites Adversarial examples are not easily detected: Bypassing ten detection methods.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Adversarial examples are not easily detected: Bypassing ten detection methods

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.595111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:67c936bf7b3c95084ac1fda7efd8cea8c457dbe5add2cf3dea032e3c00bfe4c0

Observation d0e9f695-dcd0-4105-af07-c4a89dd42499 · outbound

This paper cites Baseline Defenses for Adversarial Attacks Against Aligned Language Models.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Baseline Defenses for Adversarial Attacks Against Aligned Language Models

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-07-10T15:47:23.194167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:033d960ef1de4e13933c13c88d5d083b8820793931eaba69bdbc944ed338fb9d

Observation d18e636b-2ada-4ea8-a7c7-9bc4110bc85d · outbound

This paper cites pathway suppression.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs pathway suppression

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.590152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:d2124f2a215879f4160cc604b1ed76b9c1f63e1b812e83da58d295f7e6098798

Observation bbdb9734-898d-4eb7-9835-06164d8bc6c5 · outbound

This paper cites 1.8 in clean) despite having fewer total nodes, suggesting remaining features are hyperactivated to compensate for missing pathways.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs 1.8 in clean) despite having fewer total nodes, suggesting remaining features are hyperactivated to compensate for missing pathways

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.591831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:8122b7eeaff50f1ee6498982255538a778f8a106bf26e2a37ca44556060506b0

Observation 93585ef2-b34b-4345-a227-b24445eab099 · outbound

This paper cites Write a poem about nature.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Write a poem about nature

Reference 62

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T15:47:23.586419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:fc249aa8b34ead3f4f512f136c98bdde507031ce7e4305689bdb6792599d727a

Observation 306f5feb-03f5-4ce9-ae21-5b2b940090a0 · outbound

This paper cites This is consistent with the PCA analysis in Figure 5 showing that attack trajectories diverge most strongly at layer 15.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs This is consistent with the PCA analysis in Figure 5 showing that attack trajectories diverge most strongly at layer 15

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.588591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:900696e60ac78ad950a9adea6e4b1dc0dfe2797e144e26fd9a31b6e9c74bec27

Observation 58480b53-d3ea-45b0-85a0-7983d5ebe463 · outbound

This paper cites Tell me how to make a cake.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Tell me how to make a cake

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T15:47:23.593281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:150b9c4174ce72618070fabb28b0a85ad4b00da93b128fe4d8ac66faa92d5a92

Observation ba11b773-9795-4628-b809-d633289e3bb0 · outbound

This paper cites Tell me how to make a cake.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Tell me how to make a cake

Reference 65

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T15:47:23.598410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:ba5c3c4c8829650468c55ed655626ce0e3ce9c7bf3ce3f6b541f70cf3c73d3fa

Observation 22d6d1b7-e719-42bf-ad5c-9738d335deeb · outbound

This paper cites an unresolved cited work.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Unresolved cited work

Reference 66

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T15:47:23.583262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:b862d7243906de6703d31053d0d44a394696944807634a21aa65c9ac1739a75b

Observation 5b38be97-4696-49e6-8491-2bf4370cbcce · outbound

This paper cites an unresolved cited work.

Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-07-10T15:47:23.584852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-10T15:42:46.392593Z digest=sha256:7fa4402be72a1736b5ea8a88c3384fca12a0f3d455b53e9db528ab277af2d362

Pith citing papers

Observation 6c8cb14b-cdb1-459e-bbc6-6bdb6cde63c6 · inbound

Learning to Transmit: Volatility-Aware Predictive Communication for Energy-Efficient IoT Networks cites this paper.

Learning to Transmit: Volatility-Aware Predictive Communication for Energy-Efficient IoT Networks Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T12:21:53.334309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:21:53.334309Z digest=sha256:817655df5721cbde01bf1fa087f40c2069b9b82e87945944f7167156fbf7f862

Observation 1887b201-8937-4113-8006-45742510055d · inbound

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI cites this paper.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.188293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.702144Z digest=sha256:28e104a078629c087dbf8fbca9be26eee7628ed72e96423f2f12b5c867f25851