Pre-trained decoders consistently improve multimodal translation, while pre-trained encoders help only when visual-text alignment is strong.
In Proceedings of the 55th Annual Meeting of the Association for Computational Lin- guistics (V olume 1: Long Papers), pages 1913–1924, Vancouver, Canada
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Memory Reviving, Continuing Learning and Beyond: Evaluation of Pre-trained Encoders and Decoders for Multimodal Machine Translation
Pre-trained decoders consistently improve multimodal translation, while pre-trained encoders help only when visual-text alignment is strong.