Are Deep Models Robust against Real Distortions? A Case Study on Document Image Classification

Saifullah -; Shoaib Ahmed Siddiqui; Stefan Agne; Andreas Dengel; Sheraz Ahmed

doi:10.20944/preprints202202.0058.v1

Submitted:

01 February 2022

Posted:

03 February 2022

Read the latest preprint version here

Abstract

Deep neural networks have been extensively researched in the field of document image classification to improve classification performance and have shown excellent results. However, there is little research in this area that addresses the question of how well these models would perform in a real-world environment, where the data the models are confronted with often exhibits various types of noise or distortion. In this work, we present two separate benchmark datasets, namely RVL-CDIP-D and Tobacco3482-D, to evaluate the robustness of existing state-of-the-art document image classifiers to different types of data distortions that are commonly encountered in the real world. The proposed benchmarks are generated by inserting 21 different types of data distortions with varying severity levels into the well-known document datasets RVL-CDIP and Tobacco3482, respectively, which are then used to quantitatively evaluate the impact of the different distortion types on the performance of latest document image classifiers. In doing so, we show that while the higher accuracy models also exhibit relatively higher robustness, they still severely underperform on some specific distortions, with their classification accuracies dropping from ~90% to as low as ~40% in some cases. We also show that some of these high accuracy models perform even worse than the baseline AlexNet model in the presence of distortions, with the relative decline in their accuracy sometimes reaching as high as 300-450% that of AlexNet. The proposed robustness benchmarks are made available to the community and may aid future research in this area.

Keywords:

Document Image Classification

;

Corruption Robustness

;

Robustness to Distortions

;

Model Robustness

Subject:

Computer Science and Mathematics - Computer Vision and Graphics

Copyright: This open access article is published under a Creative Commons CC BY 4.0 license, which permit the free download, distribution, and reuse, provided that the author and preprint are cited in any reuse.

Are Deep Models Robust against Real Distortions? A Case Study on Document Image Classification

Abstract

Keywords:

Subject:

MDPI Initiatives

Important Links

Subscribe