Cross-modal attention-based multimodal deep learning framework for breast cancer diagnosis using mammography and ultrasound imaging

International Journal of Development Research

Volume: 
16
Article ID: 
31079
5 pages
Research Article

Cross-modal attention-based multimodal deep learning framework for breast cancer diagnosis using mammography and ultrasound imaging

Indu P. K., Dr. Beni, G. and Dr. Rene Dev, D.

Abstract: 

Breast cancer continues to be one of the top causes of cancer-related death for women globally, highlighting the significance of prompt and precise detection. Single-modality imaging is often used by conventional computer-aided diagnosis systems, which restricts their capacity to provide additional diagnostic data. In order to diagnose breast cancer automatically, this study suggests a multimodal deep learning framework that combines ultrasound and mammography imaging. The suggested approach learns discriminative representations from heterogeneous imaging modalities by combining an ultrasonic encoder based on Swin Transformer and a mammography encoder based on ResNet101 with a cross-modal attention fusion technique. To increase lesion visibility during preprocessing, contrast enhancement utilizing Contrast Limited Adaptive Histogram Equalization (CLAHE) was used. The classification of benign and malignant conditions was done using the fused multimodal characteristics.Utilizing publicly accessible mammography and ultrasound datasets, an experimental evaluation was carried out utilizing a pseudo-paired multimodal learning approach. The suggested model demonstrated the efficacy of multimodal fusion for breast cancer diagnosis with an accuracy of 88%, an F1-score of 0.87, and an Area Under the ROC Curve (AUC) of 0.903. The findings show that cross-attention greatly enhances diagnostic performance when convolutional and transformer-based representations are combined. Future multimodal medical imaging applications and intelligent clinical decision support systems have great potential with the suggested paradigm.

DOI: 
https://doi.org/10.37118/ijdr.31079.07.2026
Download PDF: