Pith. sign in

Boosting Standard Classification Architectures Through a Ranking Regularizer

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

We employ triplet loss as a feature embedding regularizer to boost classification performance. Standard architectures, like ResNet and Inception, are extended to support both losses with minimal hyper-parameter tuning. This promotes generality while fine-tuning pretrained networks. Triplet loss is a powerful surrogate for recently proposed embedding regularizers. Yet, it is avoided due to large batch-size requirement and high computational cost. Through our experiments, we re-assess these assumptions. During inference, our network supports both classification and embedding tasks without any computational overhead. Quantitative evaluation highlights a steady improvement on five fine-grained recognition datasets. Further evaluation on an imbalanced video dataset achieves significant improvement. Triplet loss brings feature embedding characteristics like nearest neighbor to classification models. Code available at \url{http://bit.ly/2LNYEqL}.

citation-role summary

background 1

citation-polarity summary

fields

cs.CV 1

years

2019 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Adversarial Representation Learning for Text-to-Image Matching

cs.CV · 2019-08-28 · conditional · novelty 6.0

TIMAM combines an adversarial modality discriminator, norm-softmax identification losses, and a cross-modal projection matching loss with a BERT plus bidirectional LSTM text encoder, achieving state-of-the-art text-to-image retrieval on CUHK-PEDES, Flickr30K, CUB, and Flowers.

citing papers explorer

Showing 1 of 1 citing paper.

  • Adversarial Representation Learning for Text-to-Image Matching cs.CV · 2019-08-28 · conditional · none · ref 52 · internal anchor

    TIMAM combines an adversarial modality discriminator, norm-softmax identification losses, and a cross-modal projection matching loss with a BERT plus bidirectional LSTM text encoder, achieving state-of-the-art text-to-image retrieval on CUHK-PEDES, Flickr30K, CUB, and Flowers.