SSL-CPCD: Self-supervised learning with composite pretext-class discrimination for improved generalisability in endoscopic image analysis

Abstract

Data-driven methods have shown tremendous progress in medical image analysis. In this context, deep learning-based supervised methods are widely popular. However, they require a large amount of training data and face issues in generalisability to unseen datasets that hinder clinical translation. Endoscopic imaging data is characterised by large inter- and intra-patient variability that makes these models more challenging to learn representative features for downstream tasks. Thus, despite the publicly available datasets and datasets that can be generated within hospitals, most supervised models still underperform. While self-supervised learning has addressed this problem to some extent in natural scene data, there is a considerable performance gap in the medical image domain. In this paper, we propose to explore patch-level instance-group discrimination and penalisation of inter-class variation using additive angular margin within the cosine similarity metrics. Our novel approach enables models to learn to cluster similar representations, thereby improving their ability to provide better separation between different classes. Our results demonstrate significant improvement on all metrics over the state-of-the-art (SOTA) methods on the test set from the same and diverse datasets. We evaluated our approach for classification, detection, and segmentation. SSL-CPCD attains notable Top 1 accuracy of 79.77% in ulcerative colitis classification, an 88.62% mean average precision (mAP) for detection, and an 82.32% dice similarity coefficient for segmentation tasks. These represent improvements of over 4%, 2%, and 3%, respectively, compared to the baseline architectures. We demonstrate that our method generalises better than all SOTA methods to unseen datasets, reporting over 7% improvement.

Metadata

Item Type:	Article
Authors/Creators:	Xu, Z. https://orcid.org/0000-0002-3883-3716 Rittscher, J. https://orcid.org/0000-0002-8528-8298 Ali, S. https://orcid.org/0000-0003-1313-3542
Copyright, Publisher and Additional Information:	© 2024 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works.
Keywords:	Deep learning; contrastive loss; endoscopy data; generalisation; self-supervised learning
Dates:	Published (online): 10 June 2024
Institution:	The University of Leeds
Academic Units:	The University of Leeds > Faculty of Engineering & Physical Sciences (Leeds) > School of Computing (Leeds) > Artificial Intelligence
Funding Information:	Funder Grant number Crohns and Colitis UK M2023-5 SAUBRAMANIAN
Depositing User:	Symplectic Publications
Date Deposited:	13 Jun 2024 11:05
Last Modified:	07 Aug 2024 13:50
Status:	Published online
Publisher:	Institute of Electrical and Electronics Engineers
Identification Number:	10.1109/tmi.2024.3411933
Open Archives Initiative ID (OAI ID):	oai:eprints.whiterose.ac.uk:213496

CORE (COnnecting REpositories)

SSL-CPCD: Self-supervised learning with composite pretext-class discrimination for improved generalisability in endoscopic image analysis

Abstract

Metadata

Download

Accepted Version

Export

Statistics