pith. sign in

Openai’s approach to external red teaming for ai models and systems

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

citation-role summary

background 3

citation-polarity summary

years

2026 4 2025 1

verdicts

UNVERDICTED 5

roles

background 3

polarities

background 3

representative citing papers

AVISE: Framework for Evaluating the Security of AI Systems

cs.CR · 2026-04-22 · unverdicted · novelty 6.0

AVISE provides a new framework and automated SET that identifies jailbreak vulnerabilities in language models with 92% accuracy, finding all nine tested models vulnerable to an augmented Red Queen attack.

citing papers explorer

Showing 5 of 5 citing papers.