Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:11:02.149478Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 3 inbound Pith citation observations for arXiv:2505.10670.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:11:02.149478Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:51:56.450815Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T00:31:24.634056Z
77 of 77 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c43eafac-f59d-4479-8e8e-f89c0279a822 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Artificial intelligence and the future of work: Evidence from OECD countries
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bf3d3a32-c723-4ab7-a21a-243b73e9489b · outbound
Interpretable Risk Mitigation in LLM Agent Systems Dai, Chelsea Finn, Justin Fu, Kanishka Gopalakrishnan, et al
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3801ed77-8037-4a96-b87d-1b592e49f25f · outbound
Interpretable Risk Mitigation in LLM Agent Systems Mistral 7b: Open foundation models, 2023
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 54ed5ba8-5bf9-4e48-a28b-7b79fb7685e8 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Playing repeated games with Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b7eff9b-d167-419d-ae41-da6ca8c46fe7 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Concrete Problems in AI Safety
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16702f8c-2cb7-4d44-a2c2-8387f4b0ae83 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 94b9346b-084f-4af5-a678-a12f9bd12757 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e6eafe6-a930-4240-a93e-aaa30f5aa5d6 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Emergent tool use from multi-agent autocurricula
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 351ad7f1-6ab1-403d-b7ff-ef072b69c117 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8a106779-c66d-48c7-8c48-b0f78d8abf43 · outbound
Interpretable Risk Mitigation in LLM Agent Systems On the Opportunities and Risks of Foundation Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a7adc57-b623-480d-be0c-b1dd7e69a6ee · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a338b78c-533d-4e4b-9b30-639f80356236 · outbound
Interpretable Risk Mitigation in LLM Agent Systems RT-1: Robotics Transformer for Real-World Control at Scale
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d22cc2-e43f-4ec2-aff9-61a6d878a1a6 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Playing games with gpt: What can we learn about a large language model from canonical strategic games? SSRN Electronic Journal, 2023
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fa8b734c-b330-4c95-85f8-426bab50f0df · outbound
Interpretable Risk Mitigation in LLM Agent Systems Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b60e728a-df0f-4060-9aeb-3d455d436d74 · outbound
Interpretable Risk Mitigation in LLM Agent Systems What can machine learning do? workforce implications
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e55a86e7-747f-4031-8209-67c0bab512ab · outbound
Interpretable Risk Mitigation in LLM Agent Systems Evaluating Large Language Models Trained on Code
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08b4f22d-3581-40e0-aa27-dd14aaa9041a · outbound
Interpretable Risk Mitigation in LLM Agent Systems Instigating cooperation among llm agents using adaptive information modulation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72433d46-a595-481b-a5ef-2f80a6405e69 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Sparse Autoencoders Find Highly Interpretable Features in Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eadd38c1-fbcd-4b04-a105-2ea6f1e6a5e6 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Reinforcement learning in a prisoner’s dilemma
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 24f6a202-78a8-481d-8db8-1c4b79df229d · outbound
Interpretable Risk Mitigation in LLM Agent Systems Toy Models of Superposition
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3da5c266-6d9b-45d6-8d90-16eafac12480 · outbound
Interpretable Risk Mitigation in LLM Agent Systems PoGaIN: Poisson-Gaussian Image Noise Modeling from Paired Samples
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9a93bcfe-f4cc-4a52-9e56-a716c5208000 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Not All Language Model Features Are One-Dimensionally Linear
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80ab46e3-96c8-4341-b15b-4ef484f62fc1 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Some experimental games
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f9da16f4-a983-41b4-b6d2-cccdb6ba0154 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88637004-c8d5-4677-890d-d755e3357d38 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Artificial intelligence, values, and alignment
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb688a1f-ae7d-4316-9fcd-3ddffe2c900b · outbound
Interpretable Risk Mitigation in LLM Agent Systems The Llama 3 Herd of Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aea34d7-c236-4ba2-b05c-5266d05cbe3a · outbound
Interpretable Risk Mitigation in LLM Agent Systems Measuring Massive Multitask Language Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daa69867-514e-4500-ac68-9b455cec0cf3 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Measuring mathematical problem solving with the math dataset,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccc257f3-639b-4bd7-b693-d99e2d88ec53 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Cogagent: A visual language model for gui agents
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f465e08f-8581-42e6-b51b-25e319bc463c · outbound
Interpretable Risk Mitigation in LLM Agent Systems Non-linear inference time intervention: Improving llm truthfulness
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 79c959d9-a3b3-4327-86ed-96acb55e005b · outbound
Interpretable Risk Mitigation in LLM Agent Systems Towards Reasoning in Large Language Models: A Survey
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8bf0ff6-e450-4f1c-b07c-950516506ed9 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Large language models for uavs: Current state and pathways to the future
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a7a39f6f-26fb-4c28-b196-bc2a237200ff · outbound
Interpretable Risk Mitigation in LLM Agent Systems Mixtral of Experts
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 017ef2a9-5520-4eb9-bdc4-f92641554c31 · outbound
Interpretable Risk Mitigation in LLM Agent Systems llama-3-8b-it-res (revision 53425c3), 2024
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 91339d58-f620-44d4-8eb7-b59703b0d07a · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1a302ee4-fd9f-46ae-8fab-c55bc4736872 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Martin, Hans-Theo Normann, and T
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 71a8d99b-1365-4a5b-ac7b-cce9c7fceaf8 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 53ae654e-f604-43fa-bec1-7e4bbe48d9bf · outbound
Interpretable Risk Mitigation in LLM Agent Systems Inference-Time Intervention: Eliciting Truthful Answers from a Language Model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation accf6756-7ccc-46fa-8ac0-5a95cac9c7e1 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1e40780-c265-4c8d-b588-0f4db9b277d5 · outbound
Interpretable Risk Mitigation in LLM Agent Systems The mythos of model interpretability
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4fc9c058-9b7f-4030-b591-6b05a8f26e5f · outbound
Interpretable Risk Mitigation in LLM Agent Systems Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cdffd23-9790-4415-81f6-49a0963613c8 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c86296d2-ed87-49d3-b9f2-78d6f82f7afa · outbound
Interpretable Risk Mitigation in LLM Agent Systems Linguistic regularities in continuous space word representations
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a59a7951-c2e9-43c9-a557-199380bc4da6 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Large Language Models: A Survey
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fae7cda6-218c-4394-8746-aa9f8ca079ab · outbound
Interpretable Risk Mitigation in LLM Agent Systems A strategy of win-stay, lose-shift that outperforms tit-for-tat in the prisoner’s dilemma game
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4826b429-d15b-417a-b168-ff52c95b5cbc · outbound
Interpretable Risk Mitigation in LLM Agent Systems Training language models to follow instructions with human feedback
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 014ab367-984a-479d-83c0-103ef33837a1 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Cooperation: A systematic review of how to enable agent to circumvent the prisoner’s dilemma
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7a05f4db-7afc-4148-a98b-9ef26cb1000d · outbound
Interpretable Risk Mitigation in LLM Agent Systems Generative Agents: Interactive Simulacra of Human Behavior
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de77e093-1471-462a-b0a6-a927f6729dc5 · outbound
Interpretable Risk Mitigation in LLM Agent Systems TinyClick: Single-Turn Agent for Empowering GUI Automation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a4ef3f-38c0-44df-85f6-445319704ac3 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 35196dd2-37eb-4576-a6c2-b40e2da03ea4 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Effect of private deliberation: Deception of large language models in game play
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a70cb5aa-ff92-409f-b95a-3405b51bbf27 · outbound
Interpretable Risk Mitigation in LLM Agent Systems GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5a5b1b9-7b0f-427c-973e-a741a2cd28cd · outbound
Interpretable Risk Mitigation in LLM Agent Systems A primer in BERTology: What we know about how BERT works
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb2fe07f-6a82-4fb7-bc5a-7b4cefa20efd · outbound
Interpretable Risk Mitigation in LLM Agent Systems Research priorities for robust and beneficial artificial intelligence
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7a2cf6c4-3cb3-42c1-a6c5-e33c3fd83511 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 95101970-bd83-42b3-bfc5-703b311591ce · outbound
Interpretable Risk Mitigation in LLM Agent Systems Toolformer: Language Models Can Teach Themselves to Use Tools
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f6167c9-4a5f-46af-b08b-065532a5c700 · outbound
Interpretable Risk Mitigation in LLM Agent Systems An evolutionary model of personality traits related to cooperative behavior using a large language model
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0b622b1d-44e4-43db-bc95-870dd9ca5525 · outbound
Interpretable Risk Mitigation in LLM Agent Systems A comparative analysis of the definitions of autonomous weapons systems
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 54cdc25e-dcbe-4a06-a3ca-a0b083ec66e7 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Gemma: Open Models Based on Gemini Research and Technology
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e70192c0-2fa3-4d26-9e99-25bce4521ab8 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Gemma 2: Improving Open Language Models at a Practical Size
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 625f59c6-0eef-406d-9cf0-d8c92c8ef449 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Scaling monosemanticity: Extracting interpretable features from claude 3 sonnet
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b2ebf77-d4d1-47eb-846b-c88fd4d56a1d · outbound
Interpretable Risk Mitigation in LLM Agent Systems Moral alignment for llm agents
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f523afad-184f-41ef-9a0e-57e19a11c506 · outbound
Interpretable Risk Mitigation in LLM Agent Systems LLaMA: Open and Efficient Foundation Language Models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82ac90b8-2d66-452e-9693-66c10e7d314b · outbound
Interpretable Risk Mitigation in LLM Agent Systems Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20ce1bc9-e4d8-4904-9888-9de2d34c5479 · outbound
Interpretable Risk Mitigation in LLM Agent Systems GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2942ef8-188f-4e1f-a44b-0747ba8e3471 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Mobile-agent: Autonomous multi-modal mobile device agent with visual perception,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e8c5f846-e91b-401e-80a4-0fd5d73e1194 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Dai, and Quoc V Le
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9aac0e58-4508-4dce-b103-c860ac6c28f4 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cea8ac7-27e1-49cf-b37d-d6744ff11b68 · outbound
Interpretable Risk Mitigation in LLM Agent Systems ReAct: Synergizing Reasoning and Acting in Language Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 647ac5ed-565f-4692-8bee-b21bc1834a89 · outbound
Interpretable Risk Mitigation in LLM Agent Systems AppAgent: Multimodal Agents as Smartphone Users
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33d30244-b423-4420-bcc9-fbd665b74b6d · outbound
Interpretable Risk Mitigation in LLM Agent Systems Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10bb76ee-bc2d-40f4-babf-40043502314b · outbound
Interpretable Risk Mitigation in LLM Agent Systems Representation Engineering: A Top-Down Approach to AI Transparency
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1633c79-0375-4295-bfa3-8325c644a49a · outbound
Interpretable Risk Mitigation in LLM Agent Systems You Only Look at Screens: Multimodal Chain-of-Action Agents
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d843c580-4512-4d8f-a3a6-6b41724a5ec6 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 671ed4ec-d1ad-4d99-97f8-e0d15322b52b · outbound
Interpretable Risk Mitigation in LLM Agent Systems Measuring Mathematical Problem Solving With the MATH Dataset
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b1e14ac-f124-489b-84f4-00a88d7ea70d · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 98fbd528-7341-4dfe-92d5-440307b35277 · outbound
Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b72a9458-11dc-47d1-84f0-f77d9122b3d6 · inbound
A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy Interpretable Risk Mitigation in LLM Agent Systems
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b7081ac-e99d-43fc-be2e-190be39eed65 · inbound
LLM Harms: A Taxonomy and Discussion Interpretable Risk Mitigation in LLM Agent Systems
Reference 189
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 84d058c8-51a1-4859-970e-8b0c97117aac · inbound
LLM Harms: A Taxonomy and Discussion Interpretable Risk Mitigation in LLM Agent Systems
Reference 189
Source-reported events for the cited work
Unavailable: canonical work link unavailable.