Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise

Bidur Khanal¹⁴,
Tianhong Dai¹⁶,
Binod Bhattarai¹⁶ &
…
Cristian Linte^14,15

Part of the book series: Lecture Notes in Computer Science ((LNCS,volume 15011))

Included in the following conference series:

International Conference on Medical Image Computing and Computer-Assisted Intervention

964 Accesses

Abstract

The robustness of supervised deep learning-based medical image classification is significantly undermined by label noise in the training data. Although several methods have been proposed to enhance classification performance in the presence of noisy labels, they face some challenges: 1) a struggle with class-imbalanced datasets, leading to the frequent overlooking of minority classes as noisy samples; 2) a singular focus on maximizing performance using noisy datasets, without incorporating experts-in-the-loop for actively cleaning the noisy labels. To mitigate these challenges, we propose a two-phase approach that combines Learning with Noisy Labels (LNL) and active learning. This approach not only improves the robustness of medical image classification in the presence of noisy labels but also iteratively improves the quality of the dataset by relabeling the important incorrect labels, under a limited annotation budget. Furthermore, we introduce a novel Variance of Gradients approach in the LNL phase, which complements the loss-based sample selection by also sampling under-represented examples. Using two imbalanced noisy medical classification datasets, we demonstrate that our proposed technique is superior to its predecessors at handling class imbalance by not misidentifying clean samples from minority classes as mostly noisy samples. Code available at: https://github.com/Bidur-Khanal/imbalanced-medical-active-label-cleaning.git.

B. Bhattarai and C. Linte—These authors share equal senior authorship.

This is a preview of subscription content, log in via an institution to check access.

Access this chapter

Subscribe and save

Springer+ Basic

$34.99 /Month

Get 10 units per month
Download Article/Chapter or eBook
1 Unit = 1 Article or 1 Chapter
Cancel anytime

Buy Now

Chapter: USD 29.95; Price excludes VAT (USA)

eBook: USD 99.99; Price excludes VAT (USA)

Softcover Book: USD 119.99; Price excludes VAT (USA)

Tax calculation will be finalised at checkout

Purchases are for personal use only

Institutional subscriptions

Gradient and Feature Conformity-Steered Medical Image Classification with Noisy Labels

Active label cleaning for improved dataset quality under resource constraints

Article Open access 04 March 2022

Improving Medical Image Classification in Noisy Labels Using only Self-supervised Pretraining

Notes

1.
https://challenge.isic-archive.com/landing/2019/.

References

Agarwal, C., D’souza, D., Hooker, S.: Estimating example difficulty using variance of gradients. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (2022)
Google Scholar
Bernhardt, M., Castro, D.C., Tanno, R., Schwaighofer, A., Tezcan, K.C., Monteiro, M., Bannur, S., Lungren, M.P., Nori, A., Glocker, B., et al.: Active label cleaning for improved dataset quality under resource constraints. Nature communications (2022)
Google Scholar
Budd, S., Robinson, E.C., Kainz, B.: A survey on active learning and human-in-the-loop deep learning for medical image analysis. Medical Image Analysis (2021)
Google Scholar
Cohn, D.A., Ghahramani, Z., Jordan, M.I.: Active learning with statistical models. Journal of artificial intelligence research (1996)
Google Scholar
Cui, Y., Jia, M., Lin, T.Y., Song, Y., Belongie, S.: Class-balanced loss based on effective number of samples. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition (2019)
Google Scholar
Goh, H.W., Mueller, J.: Activelab: Active learning with re-labeling by multiple annotators. In: ICLR Workshop on Trustworthy ML (2023)
Google Scholar
Han, B., Yao, Q., Yu, X., Niu, G., Xu, M., Hu, W., Tsang, I., Sugiyama, M.: Co-teaching: Robust training of deep neural networks with extremely noisy labels. Advances in neural information processing systems (2018)
Google Scholar
Irvin, J., Rajpurkar, P., Ko, M., Yu, Y., Ciurea-Ilcus, S., Chute, C., Marklund, H., Haghgoo, B., Ball, R., Shpanskaya, K., et al.: Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison. In: Proceedings of the AAAI conference on artificial intelligence (2019)
Google Scholar
Karimi, D., Dou, H., Warfield, S.K., Gholipour, A.: Deep learning with noisy labels: Exploring techniques and remedies in medical image analysis. Medical image analysis (2020)
Google Scholar
Kather, J.N., Krisam, J., Charoentong, P., Luedde, T., Herpel, E., Weis, C.A., Gaiser, T., Marx, A., Valous, N.A., Ferber, D., et al.: Predicting survival from colorectal cancer histology slides using deep learning: A retrospective multicenter study. PLoS medicine (2019)
Google Scholar
Khanal, B., Bhattarai, B., Khanal, B., Linte, C.A.: Improving medical image classification in noisy labels using only self-supervised pretraining. In: MICCAI Workshop on Data Engineering in Medical Imaging. Springer (2023)
Google Scholar
Khanal, B., Hasan, S.K., Khanal, B., Linte, C.A.: Investigating the impact of class-dependent label noise in medical image classification. In: Medical Imaging 2023: Image Processing. SPIE (2023)
Google Scholar
Kuznetsova, A., Rom, H., Alldrin, N., Uijlings, J., Krasin, I., Pont-Tuset, J., Kamali, S., Popov, S., Malloci, M., Kolesnikov, A., et al.: The open images dataset v4: Unified image classification, object detection, and visual relationship detection at scale. International Journal of Computer Vision (2020)
Google Scholar
Li, J., Cao, H., Wang, J., Liu, F., Dou, Q., Chen, G., Heng, P.A.: Learning robust classifier for imbalanced medical image dataset with noisy labels by minimizing invariant risk. In: International Conference on Medical Image Computing and Computer-Assisted Intervention. Springer (2023)
Google Scholar
Li, J., Socher, R., Hoi, S.C.: Dividemix: Learning with noisy labels as semi-supervised learning. arXiv preprint arXiv:2002.07394 (2020)
Lin, C., Mausam, M., Weld, D.: Re-active learning: Active learning with relabeling. In: Proceedings of the AAAI Conference on Artificial Intelligence (2016)
Google Scholar
Liu, J., Li, R., Sun, C.: Co-correcting: noise-tolerant medical image classification via mutual label correction. IEEE Transactions on Medical Imaging (2021)
Google Scholar
Ørting, S.N., Doyle, A., van Hilten, A., Hirth, M., Inel, O., Madan, C.R., Mavridis, P., Spiers, H., Cheplygina, V.: A survey of crowdsourcing in medical image analysis. Human Computation (2020)
Google Scholar
Rochester Institute of Technology: Research computing services (2022), https://www.rit.edu/researchcomputing/
Sener, O., Savarese, S.: Active learning for convolutional neural networks: A core-set approach. In: International Conference on Learning Representations (2018)
Google Scholar
Shin, S., Bae, H., Shin, D., Joo, W., Moon, I.C.: Loss-curvature matching for dataset selection and condensation. In: International Conference on Artificial Intelligence and Statistics. PMLR (2023)
Google Scholar
Xue, C., Yu, L., Chen, P., Dou, Q., Heng, P.A.: Robust medical image classification from noisy labeled data with global and local representation guided co-training. IEEE Transactions on Medical Imaging (2022)
Google Scholar
Zeni, M., Zhang, W., Bignotti, E., Passerini, A., Giunchiglia, F.: Fixing mislabeling by human annotators leveraging conflict resolution and prior knowledge. Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies (2019)
Google Scholar

Download references

Acknowledgments

Research reported in this publication was supported by the NIGMS Award No. R35GM128877 of the National Institutes of Health, and by OAC Award No. 1808530 and CBET Award No. 2245152, both of the National Science Foundation, and by the Aberdeen Startup Grant CF10834-10. We also acknowledge Research Computing at the Rochester Institute of Technology [19] for providing computing resources.

Author information

Authors and Affiliations

Center for Imaging Science, Rochester Institute of Technology, Rochester, NY, USA
Bidur Khanal & Cristian Linte
Biomedical Engineering, Rochester Institute of Technology, Rochester, NY, USA
Cristian Linte
University of Aberdeen, Aberdeen, UK
Tianhong Dai & Binod Bhattarai

Authors

Bidur Khanal
View author publications
You can also search for this author in PubMed Google Scholar
Tianhong Dai
View author publications
You can also search for this author in PubMed Google Scholar
Binod Bhattarai
View author publications
You can also search for this author in PubMed Google Scholar
Cristian Linte
View author publications
You can also search for this author in PubMed Google Scholar

Corresponding author

Correspondence to Bidur Khanal .

Editor information

Editors and Affiliations

Children’s National Hospital/George Washington University, Washington, DC, USA
Marius George Linguraru
The Chinese University of Hong Kong, Hong Kong, China
Qi Dou
Technical University of Denmark, Kgs Lyngby, Denmark
Aasa Feragen
Imperial College London, London, UK
Stamatia Giannarou
Imperial College London, London, UK
Ben Glocker
Universitat de Barcelona, Barcelona, Spain
Karim Lekadir
Helmholtz Munich, Technical University of Munich and King’s College London, Munich, Germany
Julia A. Schnabel

Ethics declarations

Disclosure of Interests

The authors have no competing interests to declare that are relevant to the content of this article.

1 Electronic supplementary material

Below is the link to the electronic supplementary material.

Supplementary material 1 (pdf 4785 KB)

Rights and permissions

Reprints and permissions

Copyright information

About this paper

Cite this paper

Khanal, B., Dai, T., Bhattarai, B., Linte, C. (2024). Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise. In: Linguraru, M.G., et al. Medical Image Computing and Computer Assisted Intervention – MICCAI 2024. MICCAI 2024. Lecture Notes in Computer Science, vol 15011. Springer, Cham. https://doi.org/10.1007/978-3-031-72120-5_4

Download citation

DOI: https://doi.org/10.1007/978-3-031-72120-5_4
Published: 03 October 2024
Publisher Name: Springer, Cham
Print ISBN: 978-3-031-72119-9
Online ISBN: 978-3-031-72120-5
eBook Packages: Computer ScienceComputer Science (R0)

Publish with us

Policies and ethics

Societies and partnerships

The Medical Image Computing and Computer Assisted Intervention Society (opens in a new tab)

Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise

Abstract

Access this chapter

Subscribe and save

Buy Now

Similar content being viewed by others

Gradient and Feature Conformity-Steered Medical Image Classification with Noisy Labels

Active label cleaning for improved dataset quality under resource constraints

Improving Medical Image Classification in Noisy Labels Using only Self-supervised Pretraining

Notes

References

Acknowledgments

Author information

Authors and Affiliations

Corresponding author

Editor information

Editors and Affiliations

Ethics declarations

Disclosure of Interests

1 Electronic supplementary material

Supplementary material 1 (pdf 4785 KB)

Rights and permissions

Copyright information

About this paper

Cite this paper

Download citation

Publish with us

Societies and partnerships

Subscribe and save

Buy Now

Navigation

Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise

Abstract

Access this chapter

Subscribe and save

Buy Now

Similar content being viewed by others

Gradient and Feature Conformity-Steered Medical Image Classification with Noisy Labels

Active label cleaning for improved dataset quality under resource constraints

Improving Medical Image Classification in Noisy Labels Using only Self-supervised Pretraining

Notes

References

Acknowledgments

Author information

Authors and Affiliations

Corresponding author

Editor information

Editors and Affiliations

Ethics declarations

Disclosure of Interests

1 Electronic supplementary material

Supplementary material 1 (pdf 4785 KB)

Rights and permissions

Copyright information

About this paper

Cite this paper

Download citation

Share this paper

Publish with us

Societies and partnerships

Search

Navigation