Embedding Text in Hyperbolic Spaces

Andrew M. Dai; Bhuwan Dhingra; Christopher J. Shallue; George E. Dahl; Mohammad Norouzi

arxiv: 1806.04313 · v1 · pith:OOTOXJXUnew · submitted 2018-06-12 · 💻 cs.CL · cs.LG

Embedding Text in Hyperbolic Spaces

Bhuwan Dhingra , Christopher J. Shallue , Mohammad Norouzi , Andrew M. Dai , George E. Dahl This is my paper

classification 💻 cs.CL cs.LG

keywords hyperbolicembeddingshierarchicaltextembeddinglearnlearnedwork

0 comments

read the original abstract

Natural language text exhibits hierarchical structure in a variety of respects. Ideally, we could incorporate our prior knowledge of this hierarchical structure into unsupervised learning algorithms that work on text data. Recent work by Nickel & Kiela (2017) proposed using hyperbolic instead of Euclidean embedding spaces to represent hierarchical data and demonstrated encouraging results when embedding graphs. In this work, we extend their method with a re-parameterization technique that allows us to learn hyperbolic embeddings of arbitrarily parameterized objects. We apply this framework to learn word and sentence embeddings in hyperbolic space in an unsupervised manner from text corpora. The resulting embeddings seem to encode certain intuitive notions of hierarchy, such as word-context frequency and phrase constituency. However, the implicit continuous hierarchy in the learned hyperbolic space makes interrogating the model's learned hierarchies more difficult than for models that learn explicit edges between items. The learned hyperbolic embeddings show improvements over Euclidean embeddings in some -- but not all -- downstream tasks, suggesting that hierarchical organization is more useful for some tasks than others.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Snomed2Vec: Random Walk and Poincar\'e Embeddings of a Clinical Knowledge Base for Healthcare Analytics
cs.LG 2019-07 unverdicted novelty 4.0

SNOMED-CT graph embeddings via random walks and Poincaré methods yield 5-6x better concept similarity and 6-20% better patient diagnosis prediction than prior embeddings.