An Analysis of Active Learning Strategies for Sequence Labeling TasksDownload PDFOpen Website

2008 (modified: 10 Nov 2022)EMNLP 2008Readers: Everyone
Abstract: Active learning is well-suited to many problems in natural language processing, where unlabeled data may be abundant but annotation is slow and expensive. This paper aims to shed light on the best active learning approaches for sequence labeling tasks such as information extraction and document segmentation. We survey previously used query selection strategies for sequence models, and propose several novel algorithms to address their shortcomings. We also conduct a large-scale empirical comparison using multiple corpora, which demonstrates that our proposed methods advance the state of the art.
0 Replies

Loading