An improved transformer network for skin cancer classification. (October 2022)
- Record Type:
- Journal Article
- Title:
- An improved transformer network for skin cancer classification. (October 2022)
- Main Title:
- An improved transformer network for skin cancer classification
- Authors:
- Xin, Chao
Liu, Zhifang
Zhao, Keyu
Miao, Linlin
Ma, Yizhao
Zhu, Xiaoxia
Zhou, Qiongyan
Wang, Songting
Li, Lingzhi
Yang, Feng
Xu, Suling
Chen, Haijiang - Abstract:
- Abstract: Background: Use of artificial intelligence to identify dermoscopic images has brought major breakthroughs in recent years to the early diagnosis and early treatment of skin cancer, the incidence of which is increasing year by year worldwide and poses a great threat to human health. Achievements have been made in the research of skin cancer image classification by using the deep backbone of the convolutional neural network (CNN). This approach, however, only extracts the features of small objects in the image, and cannot locate the important parts. Objectives: As a result, researchers of the paper turn to vision transformers (VIT) which has demonstrated powerful performance in traditional classification tasks. The self-attention is to improve the value of important features and suppress the features that cause noise. Specifically, an improved transformer network named SkinTrans is proposed. Innovations: To verify its efficiency, a three step procedure is followed. Firstly, a VIT network is established to verify the effectiveness of SkinTrans in skin cancer classification. Then multi-scale and overlapping sliding windows are used to serialize the image and multi-scale patch embedding is carried out which pay more attention to multi-scale features. Finally, contrastive learning is used which makes the similar data of skin cancer encode similarly so that the encoding results of different data are as different as possible. Main results: The experiment is carried outAbstract: Background: Use of artificial intelligence to identify dermoscopic images has brought major breakthroughs in recent years to the early diagnosis and early treatment of skin cancer, the incidence of which is increasing year by year worldwide and poses a great threat to human health. Achievements have been made in the research of skin cancer image classification by using the deep backbone of the convolutional neural network (CNN). This approach, however, only extracts the features of small objects in the image, and cannot locate the important parts. Objectives: As a result, researchers of the paper turn to vision transformers (VIT) which has demonstrated powerful performance in traditional classification tasks. The self-attention is to improve the value of important features and suppress the features that cause noise. Specifically, an improved transformer network named SkinTrans is proposed. Innovations: To verify its efficiency, a three step procedure is followed. Firstly, a VIT network is established to verify the effectiveness of SkinTrans in skin cancer classification. Then multi-scale and overlapping sliding windows are used to serialize the image and multi-scale patch embedding is carried out which pay more attention to multi-scale features. Finally, contrastive learning is used which makes the similar data of skin cancer encode similarly so that the encoding results of different data are as different as possible. Main results: The experiment is carried out based on two datasets, namely (1) HAM10000: a large dataset of multi-source dermatoscopic images of common skin cancers; (2)A clinical dataset of skin cancer collected by dermoscopy. The model proposed has achieved 94.3% accuracy on HAM10000 and 94.1% accuracy on our datasets, which verifies the efficiency of SkinTrans. Conclusions: The transformer network has not only achieved good results in natural language but also achieved ideal results in the field of vision, which also lays a good foundation for skin cancer classification based on multimodal data. This paper is convinced that it will be of interest to dermatologists, clinical researchers, computer scientists and researchers in other related fields, and provide greater convenience for patients. Highlights: An improved vision transformer-based model for skin cancer classification is proposed. Multi-scale and overlapping sliding windows, and multi-scale patch embedding is carried out. The contrastive learning method for skin cancer classification is applied. Label shuffling for a balanced sampling of skin cancer datasets is applied. Our proposed model is evaluated on two skin cancer datasets and achieved promising results. … (more)
- Is Part Of:
- Computers in biology and medicine. Volume 149(2022)
- Journal:
- Computers in biology and medicine
- Issue:
- Volume 149(2022)
- Issue Display:
- Volume 149, Issue 2022 (2022)
- Year:
- 2022
- Volume:
- 149
- Issue:
- 2022
- Issue Sort Value:
- 2022-0149-2022-0000
- Page Start:
- Page End:
- Publication Date:
- 2022-10
- Subjects:
- Skin cancer -- Vision transformer -- Classification -- Contrastive learning
Medicine -- Data processing -- Periodicals
Biology -- Data processing -- Periodicals
610.285 - Journal URLs:
- http://www.sciencedirect.com/science/journal/00104825/ ↗
http://www.elsevier.com/journals ↗ - DOI:
- 10.1016/j.compbiomed.2022.105939 ↗
- Languages:
- English
- ISSNs:
- 0010-4825
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 3394.880000
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 23337.xml