Pith. sign in

Paper Citation Record · LEDGER

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates

As of 21 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2607.28959.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.28959 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:32:06.668110Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2551cf87-a5b1-4484-a124-3b114d97f47c · outbound

This paper cites B Additional Results B.1 More Results Table 5: Additional cross-dataset robustness results on Pythia-1.4B.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates B Additional Results B.1 More Results Table 5: Additional cross-dataset robustness results on Pythia-1.4B

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:06.533912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:06.533912Z digest=sha256:47c9ac9ae999d98921afda993193f0bd6ca1b0c525ce1c09ec1cb84af4b7d736

Observation 3296c2b7-700a-42d3-b72a-66524f9d805d · outbound

This paper cites Transformer feed-forward layers are key-value memories.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Transformer feed-forward layers are key-value memories

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.762556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.762556Z digest=sha256:7da7961cfea5c449fffd0fce0f1a66aa8b9bedda65841c7169f2be0e5376b09d

Observation ab517157-5720-4a3c-b2b4-cab81dd2413f · outbound

This paper cites Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.929481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.929481Z digest=sha256:e4660d4390e516c90fdaaf2aec55fdd04451ef90c1bec732f7ee8d8766c15dd4

Observation 3dcae857-56cf-4d3b-b6a1-080cb97ec7d4 · outbound

This paper cites Impact of positional encoding: Clean and adversarial rademacher complexity for transformers under in-context regression.arXiv preprint arXiv:2512.09275,.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Impact of positional encoding: Clean and adversarial rademacher complexity for transformers under in-context regression.arXiv preprint arXiv:2512.09275,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.041299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.041299Z digest=sha256:8d391aa2d069030e882a7d2f22ce801a9adafa2e72831265eed605d32d123cb1

Observation 09b5869f-2af2-4ec9-b95d-8a3193b6c1b5 · outbound

This paper cites Scaling Trends in Language Model Robustness.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Scaling Trends in Language Model Robustness

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.086998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.086998Z digest=sha256:55902f6d855d67eda3ceb9b96665807493a8cab8d40e2cfc05bf5c6bdf53b9c6

Observation 848ec80b-7426-4632-bcef-803bd1d40a29 · outbound

This paper cites Scaling Laws for Neural Language Models.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Scaling Laws for Neural Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.155055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.155055Z digest=sha256:ef4af5ce53151dc0ed1dc01b3a79d4c13db819d95d396d16f96e874ca5b925ba

Observation 365eb461-c285-4511-9f18-379cb93a83d8 · outbound

This paper cites A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates A Mechanistic Understanding of Alignment Algorithms: A Case Study on DPO and Toxicity

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.215688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.215688Z digest=sha256:fada1a3b94053f2127aed1b6b969d9a03bc65ec73c8e9c892855cf1d020696aa

Observation 613df603-5f04-4fa7-96fe-b57127c83eb5 · outbound

This paper cites Towards understanding jailbreak attacks in llms: A representation space analysis.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Towards understanding jailbreak attacks in llms: A representation space analysis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.296886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.296886Z digest=sha256:4edc78e79a011a1cbae0db21ef6acdb976e853292f244042cd58ce516344e248

Observation 2f6c4c26-d4e2-410a-a026-854060ed12bf · outbound

This paper cites Alignment-constrained dynamic pruning for llms: Identifying and preserving alignment-critical circuits.arXiv preprint arXiv:2511.07482,.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Alignment-constrained dynamic pruning for llms: Identifying and preserving alignment-critical circuits.arXiv preprint arXiv:2511.07482,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.190586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.190586Z digest=sha256:35baafb89cc7e3358fe198f06563e9433f03bc9bb6b75804b700cf0b89f8bf4b

Observation 86cc0165-c5b4-41c8-9ded-d4cba665f6ca · outbound

This paper cites Attention sinks and compression valleys in llms are two sides of the same coin.arXiv preprint arXiv:2510.06477,.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Attention sinks and compression valleys in llms are two sides of the same coin.arXiv preprint arXiv:2510.06477,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.263234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.263234Z digest=sha256:cd95e857a0d6795bc514bdd576015f42927fd221137c344c37ae75884640e41d

Observation 8db00739-e0ec-4a77-90a9-3e8c425eda61 · outbound

This paper cites A general framework to enhance fine-tuning-based llm unlearning.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates A general framework to enhance fine-tuning-based llm unlearning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.352575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.352575Z digest=sha256:e21422b99f9b3d42cc820ada9ef9092eccaa6ecce76d9fab3ccf6f0b7d62ab26

Observation d02be9e3-24d2-422e-8a95-88d5772d846a · outbound

This paper cites Open Problems in Mechanistic Interpretability.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Open Problems in Mechanistic Interpretability

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.426090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.426090Z digest=sha256:e4fa35e567eb0a62236e7b7b5ca65d25383ed8e536140622a6b48b91ea92c322

Observation a8291210-311e-4da8-86bc-6484b77d4e1e · outbound

This paper cites Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Latent Adversarial Training Improves Robustness to Persistent Harmful Behaviors in LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.505938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.505938Z digest=sha256:53beab7cfd4663d7fc27156fa4559288892430870dda6558b32970aa71a50e03

Observation fb6ffc55-2fdb-4cc1-b425-57bc77fd86ea · outbound

This paper cites Layer by Layer: Uncovering Hidden Representations in Language Models.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Layer by Layer: Uncovering Hidden Representations in Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.577174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.577174Z digest=sha256:f3991dae737b9282fa94c51919357149b8b6eef4b40c882bd02e9b2ef7154e1c

Observation 2b339f4c-fdf9-465f-bbfb-7c52b605216c · outbound

This paper cites Tensor Trust: Interpretable Prompt Injection Attacks from an Online Game.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Tensor Trust: Interpretable Prompt Injection Attacks from an Online Game

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.653882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.653882Z digest=sha256:571a7c2db99d9213c99802de307fbbe119a60b7617afefc4ab1f31b80257adfc

Observation 4c8dac68-8c8c-4713-8032-17643f6529db · outbound

This paper cites Steering Language Models With Activation Engineering.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Steering Language Models With Activation Engineering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.729882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.729882Z digest=sha256:be4f2f5df716aec7cc314c531364e5aa1add31ebd4fc88f70b3d1ed831e7822b

Observation b46aaab2-d5e1-44f5-83a5-5aed2fab45e8 · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.802460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.802460Z digest=sha256:2f6370834a4d919c7b9cdcd1607a470471b61e0127de008a45f5f6787692497c

Observation d34c49ad-2ba7-4c9d-b8f7-0132ba070559 · outbound

This paper cites Fast is better than free: Revisiting adversarial training.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Fast is better than free: Revisiting adversarial training

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.877768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.877768Z digest=sha256:ab319e9dff27c6b4862119998c51b5eb411b834273d1a4a68f4165ece8f8b5d1

Observation 35456b1c-74db-4559-a505-52e9e6749cf2 · outbound

This paper cites Qwen2.5 Technical Report.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Qwen2.5 Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.969389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.969389Z digest=sha256:6e2344f3b3ec86b857fe524937dd2609f2d2b6b1112c6f546263fc33b537b861

Observation cf703032-cc22-4496-aa42-6183a8021c9c · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:06.059462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:06.059462Z digest=sha256:ebfbc0b4b3f853a4be290b90351c323bb0c2a09e555c2f43fd4f05c56677255e

Observation ad144c06-cf9a-4a3f-bdb6-68b24ef6127c · outbound

This paper cites Adversarial Training: A Survey.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Adversarial Training: A Survey

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:06.164646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:06.164646Z digest=sha256:ed7db5a9da9b21d3dd854866425e1f6f00348428f4fa7d0fe14c40d537ce28dd

Observation 7b1d80d9-b084-414a-9a8f-51ecc7ee23d1 · outbound

This paper cites On Prompt-Driven Safeguarding for Large Language Models.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates On Prompt-Driven Safeguarding for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:06.278220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:06.278220Z digest=sha256:e2cdebd7aaa632b04ddf12a97adf920db3c6abd41422612869a5b80cb40ff6f9

Observation dc43658d-ade5-4f40-ac50-deb247ede3c3 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Representation Engineering: A Top-Down Approach to AI Transparency

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:06.408696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:06.408696Z digest=sha256:3279c6813cb29fdd37abe6ab200e6ae58aeb546e34117ba95ae64464454ea203

Observation 0e9accf4-b6e4-42d9-80ee-6f69908a6df3 · outbound

This paper cites Hence ∥(I−P S)rt∥ ≤1 m0 Mt.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Hence ∥(I−P S)rt∥ ≤1 m0 Mt

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:06.668110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:06.668110Z digest=sha256:4127e186b80417da44465f68f6fb79f0f8acf210ed0e80d741bfac4149626d6c

Observation 7f9d2b97-93d2-4982-bbd5-09161d5f1032 · outbound

This paper cites Pruning Convolutional Neural Networks for Resource Efficient Inference.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Pruning Convolutional Neural Networks for Resource Efficient Inference

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.848482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.848482Z digest=sha256:b7717bdc1f797eb265e4895b7eeaa28f2c3dace41353124abe81ef69d7764853

Observation 55167202-7ded-402f-a8cb-7e661079240d · outbound

This paper cites Towards Deep Learning Models Resistant to Adversarial Attacks.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Towards Deep Learning Models Resistant to Adversarial Attacks

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.443242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.443242Z digest=sha256:5cbe3541c7032a82b99a3e4a848e18a420c1a9168c61cf52074097a8bf9ab1cf

Observation 43a7b00d-903e-46e1-96cd-43356c8c70fb · outbound

This paper cites The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates The Geometry of Truth: Emergent Linear Structure in Large Language Model Representations of True/False Datasets

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.612787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.612787Z digest=sha256:7c000eab859b7d8db4c2b04eafe99be0cf69e722eacad3fd4d2be6df973ee7dc

Observation 64ff00e6-bf42-4af4-a79f-c8ca2466c3d6 · outbound

This paper cites Softmax is 1/2-lipschitz: A tight bound across all ℓp norms.arXiv preprint arXiv:2510.23012,.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Softmax is 1/2-lipschitz: A tight bound across all ℓp norms.arXiv preprint arXiv:2510.23012,

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:05.019151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:05.019151Z digest=sha256:aa9e06c0d3b5ac2eecd9d17933f7f5b86947ea18ac20f56fe947e72070671977

Observation 1eed1946-ed36-448a-8bc9-922f65ba0adf · outbound

This paper cites Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.229149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.229149Z digest=sha256:f4e0d85bf73387a1a1b69a0b019e7fad2aa4dc3d99688763a78fa1dbfc8dc281

Observation 9bc1f7c2-9e55-40b0-adbf-cf44297a0133 · outbound

This paper cites The Llama 3 Herd of Models.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates The Llama 3 Herd of Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.820766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.820766Z digest=sha256:8164773ba688be64c71227a8eff33df95970bc71d5601491096e0904aa488781

Observation 47f5e25e-dfc7-4f77-a259-9fca08aa7157 · outbound

This paper cites Toy Models of Superposition.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Toy Models of Superposition

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.699008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.699008Z digest=sha256:744244ebff38ad74b5bb8167870f06b057ea0066c6cf683dbf092c72b5f25902

Observation 59d6cbe8-68ab-46d9-8c3e-f35e0b6373a3 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:04.732124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:04.732124Z digest=sha256:2bd18a6bd6f817810704fa6da5031d44c92e588193731050338c865abafae946

Observation a624ce2a-ef17-492b-88e7-954ac420c852 · outbound

This paper cites Mixat: Combining continuous and discrete adversarial training for llms.arXiv preprint arXiv:2505.16947,.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Mixat: Combining continuous and discrete adversarial training for llms.arXiv preprint arXiv:2505.16947,

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.608483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.608483Z digest=sha256:ef381f9a97d8c9e653b94aaf846fa4017818d01b33be0dd6dd5b59ffc4e0b783

Observation 8ea2f2d8-3de5-4382-8e84-a4fd76f9be3c · outbound

This paper cites Towards understanding safety alignment: A mechanistic perspective from safety neurons.arXiv preprint arXiv:2406.14144,.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Towards understanding safety alignment: A mechanistic perspective from safety neurons.arXiv preprint arXiv:2406.14144,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.423049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.423049Z digest=sha256:784190f22f53b1e6c171a54f3af44d4f7173c12346137e45001f6df91d70497b

Observation 9c23aac0-fd84-4bc8-8368-9b8bd121bffc · outbound

This paper cites Defending Against Unforeseen Failure Modes with Latent Adversarial Training.

Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates Defending Against Unforeseen Failure Modes with Latent Adversarial Training

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-03T16:32:03.312558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:32:03.312558Z digest=sha256:c6a24a28cd8f2c0c10270816dff49ca85e1f19fe6c309c1c072ee9c9234b1e2a

Pith citing papers

No inbound Pith citation observations are available.