A modular network fuses truncated YOLOv5 features with Sentence-BERT task tokens via a one-layer transformer to output task-conditioned saliency maps on a four-task eye-tracking set.
Title resolution pending
1 Pith paper cite this work, alongside 4 external citations. Polarity classification is still indexing.
1
Pith paper citing it
4
external citations · external index
fields
cs.CV 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
TDSal: Task-Based Top-Down Saliency Prediction Model
A modular network fuses truncated YOLOv5 features with Sentence-BERT task tokens via a one-layer transformer to output task-conditioned saliency maps on a four-task eye-tracking set.