Abstractive Summarization of Reddit Posts with Multi-level Memory Networks

Byeongchang Kim; Gunhee Kim; Hyunwoo Kim

arxiv: 1811.00783 · v2 · pith:HVVV5N62new · submitted 2018-11-02 · 💻 cs.CL

Abstractive Summarization of Reddit Posts with Multi-level Memory Networks

Byeongchang Kim , Hyunwoo Kim , Gunhee Kim This is my paper

classification 💻 cs.CL

keywords abstractivedatasetredditsummarizationtextmemorymulti-levelposts

0 comments

read the original abstract

We address the problem of abstractive summarization in two directions: proposing a novel dataset and a new model. First, we collect Reddit TIFU dataset, consisting of 120K posts from the online discussion forum Reddit. We use such informal crowd-generated posts as text source, in contrast with existing datasets that mostly use formal documents as source such as news articles. Thus, our dataset could less suffer from some biases that key sentences usually locate at the beginning of the text and favorable summary candidates are already inside the text in similar forms. Second, we propose a novel abstractive summarization model named multi-level memory networks (MMN), equipped with multi-level memory to store the information of text from different levels of abstraction. With quantitative evaluation and user studies via Amazon Mechanical Turk, we show the Reddit TIFU dataset is highly abstractive and the MMN outperforms the state-of-the-art summarization models.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Enriching and Controlling Global Semantics for Text Summarization
cs.CL 2021-09 unverdicted novelty 5.0

A normalizing-flow neural topic model plus control mechanism are added to Transformer summarizers to supply and regulate global semantics, with reported gains over prior models on five benchmarks.