pith. machine review for the scientific record. sign in

arxiv: 1612.04732 · v1 · submitted 2016-12-14 · 💻 cs.CL

Recognition: unknown

Multilingual Word Embeddings using Multigraphs

Authors on Pith no claims yet
classification 💻 cs.CL
keywords multilingualembeddingssemanticmodelssimilarityunsupervisedwordaccuracy
0
0 comments X
read the original abstract

We present a family of neural-network--inspired models for computing continuous word representations, specifically designed to exploit both monolingual and multilingual text. This framework allows us to perform unsupervised training of embeddings that exhibit higher accuracy on syntactic and semantic compositionality, as well as multilingual semantic similarity, compared to previous models trained in an unsupervised fashion. We also show that such multilingual embeddings, optimized for semantic similarity, can improve the performance of statistical machine translation with respect to how it handles words not present in the parallel data.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.