Harmbench: a standardized evaluation framework for automated red teaming and robust refusal,

· 2024

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

browse 1 citing papers

representative citing papers

ARMOR 2025: A Military-Aligned Benchmark for Evaluating Large Language Model Safety Beyond Civilian Contexts

cs.AI · 2026-04-30 · unverdicted · novelty 7.0

ARMOR 2025 is a new benchmark with 519 doctrinally grounded prompts across a 12-category OODA-based taxonomy that reveals critical safety gaps in 21 commercial LLMs for military applications.

citing papers explorer

Showing 1 of 1 citing paper.

ARMOR 2025: A Military-Aligned Benchmark for Evaluating Large Language Model Safety Beyond Civilian Contexts cs.AI · 2026-04-30 · unverdicted · none · ref 16
ARMOR 2025 is a new benchmark with 519 doctrinally grounded prompts across a 12-category OODA-based taxonomy that reveals critical safety gaps in 21 commercial LLMs for military applications.

Harmbench: a standardized evaluation framework for automated red teaming and robust refusal,

fields

years

verdicts

representative citing papers

citing papers explorer