← back to paper
arxiv: 2508.20333 · 2 revisions
Poison Once, Refuse Forever: Weaponizing Alignment for Injecting Bias in LLMs