Pith. sign in

MART: improving LLM safety with multi- round automatic red-teaming

9 Pith papers cite this work, alongside 4 external citations. Polarity classification is still indexing.

9 Pith papers citing it
4 external citations · external index

citation-role summary

background 2 method 1

citation-polarity summary

polarities

background 3

representative citing papers

Adaptive Instruction Composition for Automated LLM Red-Teaming

cs.CR · 2026-04-22 · unverdicted · novelty 7.0

Adaptive Instruction Composition uses a neural contextual bandit with RL to adaptively combine crowdsourced texts, generating more effective and diverse LLM jailbreaks than random or prior adaptive methods on Harmbench.

citing papers explorer

Showing 9 of 9 citing papers.