A backdoor attack that uses language as the trigger works on specific tasks, but the claimed task-agnostic generalization (BadLingual) is only demonstrated in a few settings and is contradicted by many of the paper's own experiments.
https://gemini.google.com/app
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CR 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
BadLingual: A Novel Lingual-Backdoor Attack against Large Language Models
A backdoor attack that uses language as the trigger works on specific tasks, but the claimed task-agnostic generalization (BadLingual) is only demonstrated in a few settings and is contradicted by many of the paper's own experiments.