Pith. sign in

REVIEW 2 cited by

An Ensemble Method to Produce High-Quality Word Embeddings (2016)

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1604.01692 v2 pith:6W4XIZE3 submitted 2016-04-06 cs.CL

classification cs.CL
keywords embeddingsensemblemethodwordsachieveapproachbestcombines
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
abstract

A currently successful approach to computational semantics is to represent words as embeddings in a machine-learned vector space. We present an ensemble method that combines embeddings produced by GloVe (Pennington et al., 2014) and word2vec (Mikolov et al., 2013) with structured knowledge from the semantic networks ConceptNet (Speer and Havasi, 2012) and PPDB (Ganitkevitch et al., 2013), merging their information into a common representation with a large, multilingual vocabulary. The embeddings it produces achieve state-of-the-art performance on many word-similarity evaluations. Its score of $\rho = .596$ on an evaluation of rare words (Luong et al., 2013) is 16% higher than the previous best known system.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Modeling Engagement Dynamics of Online Discussions using Relativistic Gravitational Theory

    cs.SI 2019-08 conditional novelty 6.0 of 10

    A gravity-inspired neural network predicts which user clusters engage next in Reddit discussions and forecasts comment growth rate, reporting gains over several baselines.

  2. Into the Battlefield: Quantifying and Modeling Intra-community Conflicts in Online Discussion

    cs.SI 2019-09 conditional novelty 5.0 of 10

    A continuous target-dependent sentiment difference score measures conflict in online discussion, and machine learning models predict news and user-level conflict with AUC up to 0.89.

Pith tools