Security tensors, learned input perturbations, transfer a language model's text-based safety behavior to visual inputs, improving harmful-image rejection in LVLMs while largely preserving benign performance.
Dress: Instructing large vision-language models to align and interact with humans via natural language feedback
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Security Tensors as a Cross-Modal Bridge: Extending Text-Aligned Safety to Vision in LVLM
Security tensors, learned input perturbations, transfer a language model's text-based safety behavior to visual inputs, improving harmful-image rejection in LVLMs while largely preserving benign performance.