From News to Medical: Cross-domain Discourse Segmentation

2019-04-14

Elisa Ferracane, Titan Page, Junyi Jessy Li, Katrin Erk

arXiv_CL

arXiv_CL Segmentation

Abstract
Abstract (translated by Google)
URL
PDF

Abstract

The first step in discourse analysis involves dividing a text into segments. We annotate the first high-quality small-scale medical corpus in English with discourse segments and analyze how well news-trained segmenters perform on this domain. While we expectedly find a drop in performance, the nature of the segmentation errors suggests some problems can be addressed earlier in the pipeline, while others would require expanding the corpus to a trainable size to learn the nuances of the medical domain.

Abstract (translated by Google)

URL

http://arxiv.org/abs/1904.06682

PDF

http://arxiv.org/pdf/1904.06682

From News to Medical: Cross-domain Discourse Segmentation

Abstract

Abstract (translated by Google)

URL

PDF

Comments