pith. sign in

arxiv: 1904.06682 · v1 · pith:LOQ3E6Z6new · submitted 2019-04-14 · 💻 cs.CL

From News to Medical: Cross-domain Discourse Segmentation

classification 💻 cs.CL
keywords discoursemedicalcorpusdomainfirstsegmentationsegmentswhile
0
0 comments X
read the original abstract

The first step in discourse analysis involves dividing a text into segments. We annotate the first high-quality small-scale medical corpus in English with discourse segments and analyze how well news-trained segmenters perform on this domain. While we expectedly find a drop in performance, the nature of the segmentation errors suggests some problems can be addressed earlier in the pipeline, while others would require expanding the corpus to a trainable size to learn the nuances of the medical domain.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.