SalGAN: Visual Saliency Prediction with Generative Adversarial Networks

Cristian Canton Ferrer; Elisa Sayrol; Jordi Torres; Junting Pan; Kevin McGuinness; Noel E. O'Connor; Xavier Giro-i-Nieto

arxiv: 1701.01081 · v3 · pith:NPIDZ47Onew · submitted 2017-01-04 · 💻 cs.CV

SalGAN: Visual Saliency Prediction with Generative Adversarial Networks

Junting Pan , Cristian Canton Ferrer , Kevin McGuinness , Noel E. O'Connor , Jordi Torres , Elisa Sayrol , Xavier Giro-i-Nieto This is my paper

classification 💻 cs.CV

keywords saliencyadversarialnetworkpredictiontrainedbinarygenerativeloss

0 comments

read the original abstract

We introduce SalGAN, a deep convolutional neural network for visual saliency prediction trained with adversarial examples. The first stage of the network consists of a generator model whose weights are learned by back-propagation computed from a binary cross entropy (BCE) loss over downsampled versions of the saliency maps. The resulting prediction is processed by a discriminator network trained to solve a binary classification task between the saliency maps generated by the generative stage and the ground truth ones. Our experiments show how adversarial training allows reaching state-of-the-art performance across different metrics when combined with a widely-used loss function like BCE. Our results can be reproduced with the source code and trained models available at https://imatge-upc.github.io/saliency-salgan-2017/.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Data-centric Design of Learning-based Surgical Gaze Perception Models in Multi-Task Simulation
cs.RO 2026-02 unverdicted novelty 4.0

Introduces a multi-task surgical gaze dataset comparing active execution versus passive viewing and novice versus intermediate expertise, showing passive novice labels approximate intermediate active attention with li...