Pith. sign in

Title resolution pending

3 Pith papers cite this work, alongside 20 external citations. Polarity classification is still indexing.

3 Pith papers citing it
20 external citations · external index

citation-role summary

background 1

citation-polarity summary

fields

cs.AI 3

verdicts

UNVERDICTED 3

roles

background 1

polarities

background 1

representative citing papers

Provably Secure Agent Guardrail

cs.AI · 2026-05-28 · unverdicted · novelty 6.0

Introduces ePCA framework using neural-symbolic isolation to force agents to formalize intentions as logical constraints, claiming zero attack success and false positive rates in tested scenarios.

citing papers explorer

Showing 3 of 3 citing papers.

  • Provably Secure Agent Guardrail cs.AI · 2026-05-28 · unverdicted · none · ref 19

    Introduces ePCA framework using neural-symbolic isolation to force agents to formalize intentions as logical constraints, claiming zero attack success and false positive rates in tested scenarios.

  • AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions cs.AI · 2024-08-23 · unverdicted · none · ref 248

    The paper introduces a taxonomy of AI safety for LLMs organized into Trustworthy AI, Responsible AI, and Safe AI perspectives, accompanied by a review of state-of-the-art methods, challenges, and future directions.

  • Robust AI Security and Alignment: A Sisyphean Endeavor? cs.AI · 2025-12-10 · unverdicted · none · ref 3

    AI security and alignment cannot achieve full robustness because any sufficiently powerful AI inherits incompleteness-style limitations from formal systems.