Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T20:22:29.726790Z
Paper Citation Record · LEDGER
As of 2 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 4 inbound Pith citation observations for arXiv:2512.20677.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T20:22:29.726790Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-02T06:30:47.504484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T17:36:18.728469Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T15:28:33.588948Z
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 246968f4-0ce8-41ff-baf0-d63653468f6e · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models A General Language Assistant as a Laboratory for Alignment
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation af39c7d1-2d5d-46ed-a085-41c01451a666 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models LLM-Safety Evaluations Lack Robustness
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation ee97a497-fdd1-4c7f-b2cf-3da30878298e · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Measuring Massive Multitask Language Understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 7560045d-ed96-4102-be1d-3db6c289e96d · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Adversarial Machine Learning at Scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 580e4272-b477-4297-b2f3-b2c0e9b0c0e9 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Paul Röttger, Hannah Rose Kirk, Bertie Vidgen, Giuseppe Attanasio, Federico Bianchi, and Dirk Hovy
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 04cb83cd-e566-4c35-a53d-72324df03873 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Survey of Vulnerabilities in Large Language Models Revealed by Adversarial Attacks
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation bc2eb868-8ecd-46a3-9d65-9f2663b84adc · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models PandaGuard: Systematic Evaluation of LLM Safety against Jailbreaking Attacks
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation d0c0a8d8-acc8-431d-a17c-b81bce96d53b · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Mujiangshan Wang, Yuqing Lin, Shiying Wang, and Meiyu Wang
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation c678e7ef-4abf-4fc9-bb57-f588bb30f4c6 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Mujiangshan Wang, Shuhao Xu, Jincheng Jiang, Dong Xiang, and Sun-Yuan Hsieh
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 17ad310a-1dde-4d4c-abe3-e439c04b3187 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Ethical and social risks of harm from Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 319ce88b-2abd-4e41-b083-0c7b5201173b · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 4152c7c1-9b26-464d-ba8f-35c853a6d0e4 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 5f81880f-7696-4e58-9b60-7b9d15a8c3a0 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 6b0ce669-c647-4e61-89b2-22f7f01fd34a · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models "" 24returntemplate Listing 1: Data exfiltration attack generation template A.2 Sandbagging Templates 1defgenerate_sandbagging_template( task_domain): 2template = f
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 3ecf64d1-1a65-4a2f-8ec6-73a8d7709b81 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation d4a66ce0-9bf3-47d4-beaf-be715c0d6fc1 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 51a8e40d-4412-4776-aace-6f9d3cf47f95 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation f9f49c49-dac6-4ffe-88f9-73f5b423d14a · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models " 18returntemplate Listing 2: Sandbagging attack generation template A.3 Inappropriate Tool Use Templates 1defgenerate_tool_misuse_template(tool_name, usage_context): 2template = f
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation a0342f47-783d-4eef-b45b-e25f06e9679b · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation fed0c238-e3bc-47e8-ba31-efc09560a283 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 50445822-c4d7-4d70-8c18-14b89c082b0d · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 0035dc87-f094-41d6-9b63-dc1c6b13028d · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation f21d96f9-a731-429b-ab46-f68e5fcba442 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation e5a2d740-7883-4945-bb99-eed0be85fbc9 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 3a215295-4c8c-4d34-a6a9-ed4863ed8a08 · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 47ced916-de59-45d9-aaa5-16391b75206a · outbound
Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models your-api-key-here
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 8bdc3a9f-8905-4235-95bc-bcdbf89d1307 · inbound
EVLA: An Electro-Aware Multimodal Assistant for Physically-Grounded Driving Reasoning and Control Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 2de3765d-1d6c-4f66-bac5-1998e88b5895 · inbound
A3M: Adaptive, Adversarial and Multi-Objective Learning for Strategic Bidding in Repeated Auctions Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation c48e850b-fcdb-496f-885d-6be9914c966d · inbound
Beyond Skepticism: Evaluating LLMs Pedagogical Intent Reasoning with the Adaptive Pedagogical Vigilance Framework Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.
Observation 58dfb0a0-46e7-4d5f-b52f-68758983d879 · inbound
FedCausal-Dyn: A Causal-Dynamic Paradigm for Federated Learning under Dynamic Feature Drift Learning-Based Automated Adversarial Red-Teaming for Robustness Evaluation of Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.