pith. machine review for the scientific record. sign in

arxiv: 1901.10444 · v1 · submitted 2019-01-29 · 💻 cs.CL

Recognition: unknown

No Training Required: Exploring Random Encoders for Sentence Classification

Authors on Pith no claims yet
classification 💻 cs.CL
keywords sentenceembeddingsrandomclassificationtrainingturnsappropriatebaselines
0
0 comments X
read the original abstract

We explore various methods for computing sentence representations from pre-trained word embeddings without any training, i.e., using nothing but random parameterizations. Our aim is to put sentence embeddings on more solid footing by 1) looking at how much modern sentence embeddings gain over random methods---as it turns out, surprisingly little; and by 2) providing the field with more appropriate baselines going forward---which are, as it turns out, quite strong. We also make important observations about proper experimental protocol for sentence classification evaluation, together with recommendations for future research.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Emergent Manifold Separability during Reasoning in Large Language Models

    cs.LG 2026-02 unverdicted novelty 6.0

    Reasoning in LLMs produces a transient geometric pulse in which concept manifolds untangle into linearly separable subspaces immediately before computation and compress afterward.