TNN-C-CCA, a two-stage pipeline of Cluster-CCA embeddings refined by a deep triplet network with cosine triplet loss, achieves state-of-the-art MAP on VEGAS and MV-10K audio-visual retrieval benchmarks.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.MM 1years
2019 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Deep Triplet Neural Networks with Cluster-CCA for Audio-Visual Cross-modal Retrieval
TNN-C-CCA, a two-stage pipeline of Cluster-CCA embeddings refined by a deep triplet network with cosine triplet loss, achieves state-of-the-art MAP on VEGAS and MV-10K audio-visual retrieval benchmarks.