Pith. sign in

REVIEW

Unity in Diversity: Learning Distributed Heterogeneous Sentence Representation for Extractive Summarization

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1912.11688 v1 pith:4BFJK5WE submitted 2019-12-25 cs.CL cs.IRcs.LG

classification cs.CLcs.IRcs.LG
keywords sentenceextractivefeaturessummarycompositionalrepresentationsemanticsentences
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Automated multi-document extractive text summarization is a widely studied research problem in the field of natural language understanding. Such extractive mechanisms compute in some form the worthiness of a sentence to be included into the summary. While the conventional approaches rely on human crafted document-independent features to generate a summary, we develop a data-driven novel summary system called HNet, which exploits the various semantic and compositional aspects latent in a sentence to capture document independent features. The network learns sentence representation in a way that, salient sentences are closer in the vector space than non-salient sentences. This semantic and compositional feature vector is then concatenated with the document-dependent features for sentence ranking. Experiments on the DUC benchmark datasets (DUC-2001, DUC-2002 and DUC-2004) indicate that our model shows significant performance gain of around 1.5-2 points in terms of ROUGE score compared with the state-of-the-art baselines.

Discussion (0). Sign in to comment.

Pith tools