Revisiting Deep Feature Reconstruction for Logical and Structural Industrial Anomaly Detection

Sukanya Patra; Souhaib Ben Taieb

Revisiting Deep Feature Reconstruction for Logical and Structural Industrial Anomaly Detection

Sukanya Patra, Souhaib Ben Taieb

Published: 07 Oct 2024, Last Modified: 07 Oct 2024Accepted by TMLREveryoneRevisionsBibTeXCC BY 4.0

Abstract: Industrial anomaly detection is crucial for quality control and predictive maintenance, but it presents challenges due to limited training data, diverse anomaly types, and external factors that alter object appearances. Existing methods commonly detect structural anomalies, such as dents and scratches, by leveraging multi-scale features from image patches extracted through deep pre-trained networks. However, significant memory and computational demands often limit their practical application. Additionally, detecting logical anomalies—such as images with missing or excess elements—requires an understanding of spatial relationships that traditional patch-based methods fail to capture. In this work, we address these limitations by focusing on Deep Feature Reconstruction (DFR), a memory- and compute-efficient approach for detecting structural anomalies. We further enhance DFR into a unified framework, called ULSAD, which is capable of detecting both structural and logical anomalies. Specifically, we refine the DFR training objective to improve performance in structural anomaly detection, while introducing an attention-based loss mechanism using a global autoencoder-like network to handle logical anomaly detection. Our empirical evaluation across five benchmark datasets demonstrates the performance of ULSAD in detecting and localizing both structural and logical anomalies, outperforming eight state-of-the-art methods. An extensive ablation study further highlights the contribution of each component to the overall performance improvement. Our code is available at https://github.com/sukanyapatra1997/ULSAD-2024.git.

Submission Length: Regular submission (no more than 12 pages of main content)

Changes Since Last Submission: Camera Ready Version

Code: https://github.com/sukanyapatra1997/ULSAD-2024.git

Assigned Action Editor: ~Yan_Liu1

Submission Number: 2579

Loading