Crowdsourced Corpus with Entity Salience AnnotationsDownload PDFOpen Website

2016 (modified: 06 Nov 2022)LREC 2016Readers: Everyone
Abstract: In this paper, we present a crowdsourced dataset which adds entity salience (importance) annotations to the Reuters-128 dataset, which is subset of Reuters-21578. The dataset is distributed under a free license and publish in the NLP Interchange Format, which fosters interoperability and re-use. We show the potential of the dataset on the task of learning an entity salience classifier and report on the results from several experiments.
0 Replies

Loading