MAGE uses an agentic shadow memory to proactively detect and mitigate long-horizon threats in LLM agents by distilling safety context and assessing action risks before execution.
Title resolution pending
2 Pith papers cite this work. Polarity classification is still indexing.
fields
cs.CR 2years
2026 2representative citing papers
Alignment contracts define scope, allowed effects, budgets and disclosure rules as safety properties over finite effect traces, with decidable admissibility, refinement rules, and Lean-verified soundness under an observability assumption.
citing papers explorer
-
MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory
MAGE uses an agentic shadow memory to proactively detect and mitigate long-horizon threats in LLM agents by distilling safety context and assessing action risks before execution.
-
Alignment Contracts for Agentic Security Systems
Alignment contracts define scope, allowed effects, budgets and disclosure rules as safety properties over finite effect traces, with decidable admissibility, refinement rules, and Lean-verified soundness under an observability assumption.