Deep k-NN for Noisy Labels

Dara Bahri; Heinrich Jiang; Maya Gupta

Deep k-NN for Noisy Labels

Dara Bahri, Heinrich Jiang, Maya Gupta

25 Sept 2019 (modified: 22 Jun 2025)ICLR 2020 Conference Blind SubmissionReaders: Everyone

Abstract: Modern machine learning models are often trained on examples with noisy labels that hurt performance and are hard to identify. In this paper, we provide an empirical study showing that a simple $k$-nearest neighbor-based filtering approach on the logit layer of a preliminary model can remove mislabeled training data and produce more accurate models than some recently proposed methods. We also provide new statistical guarantees into its efficacy.

Community Implementations: [![CatalyzeX](/images/catalyzex_icon.svg) 2 code implementations](https://www.catalyzex.com/paper/deep-k-nn-for-noisy-labels/code)

Original Pdf: pdf

6 Replies

Loading