REVIEW 2 cited by
Polyglot: Distributed Word Representations for Multilingual NLP
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Distributed word representations (word embeddings) have recently contributed to competitive performance in language modeling and several NLP tasks. In this work, we train word embeddings for more than 100 languages using their corresponding Wikipedias. We quantitatively demonstrate the utility of our word embeddings by using them as the sole features for training a part of speech tagger for a subset of these languages. We find their performance to be competitive with near state-of-art methods in English, Danish and Swedish. Moreover, we investigate the semantic features captured by these embeddings through the proximity of word groupings. We will release these embeddings publicly to help researchers in the development and enhancement of multilingual applications.
Forward citations
Cited by 2 Pith papers
-
Classifying single-qubit noise using machine learning
Supervised classifiers distinguish coherent from stochastic single-qubit noise on GST data, with near-perfect accuracy after feature engineering and margin-based robustness to sampling noise.
-
Hierarchical Pointer Net Parsing
A hierarchical pointer-network decoder that conditions on parent and sibling states improves discourse parsing relation F1 to 82.77 and gives marginal dependency parsing gains.
Discussion (0). Continue with ORCID to comment.