Showing 1–1 of 1 results for author: Page, T
-
From News to Medical: Cross-domain Discourse Segmentation
Authors:
Elisa Ferracane,
Titan Page,
Junyi Jessy Li,
Katrin Erk
Abstract:
The first step in discourse analysis involves dividing a text into segments. We annotate the first high-quality small-scale medical corpus in English with discourse segments and analyze how well news-trained segmenters perform on this domain. While we expectedly find a drop in performance, the nature of the segmentation errors suggests some problems can be addressed earlier in the pipeline, while…
▽ More
The first step in discourse analysis involves dividing a text into segments. We annotate the first high-quality small-scale medical corpus in English with discourse segments and analyze how well news-trained segmenters perform on this domain. While we expectedly find a drop in performance, the nature of the segmentation errors suggests some problems can be addressed earlier in the pipeline, while others would require expanding the corpus to a trainable size to learn the nuances of the medical domain.
△ Less
Submitted 14 April, 2019;
originally announced April 2019.