Abstract
Breast cancer, a leading cause of cancer-related deaths in women, presents a growing challenge in medical diagnostics. Despite the effectiveness of mammography and ultrasound, the ambiguity in non-invasive scans often necessitates invasive procedures. Our primary goal was to create an AI model that could predict breast cancer with high negative predictive value and reduce unnecessary procedures. This study introduces the Retrieval-Augmented Medical Diagnosis System (RAMDS) for breast cancer, a novel approach combining an AI model with a retrieval-augmented mechanism to enhance diagnostic accuracy and explainability. The RAMDS employs a pretrained ResNet 34 model, fine-tuned on breast ultrasound image datasets from four countries. It integrates a similarity-based weighted adjustment mechanism to compare new cases with historical diagnoses. It’s like having an experienced doctor who remembers every case they’ve ever seen and uses that knowledge to make better decisions. RAMDS improved sensitivity by 11%, and negative predictive value by 9% when compared to the base model. Notably, the RAMDS improves explainability by linking AI predictions to similar historical cases, aligning with the medical community’s interest in transparent and understandable AI decisions.
A unique feature of this system is its adaptability to varied imaging contexts without retraining, addressing the challenge of dataset variability across medical institutions. In conclusion, the RAMDS offers a significant advancement in breast cancer diagnosis, combining enhanced accuracy, explainability, and adaptability. It holds promise for clinical application, though further research is needed to optimize its performance and integrate multi-modal data.
Competing Interest Statement
The authors have declared no competing interest.
Funding Statement
No funding was received
Author Declarations
I confirm all relevant ethical guidelines have been followed, and any necessary IRB and/or ethics committee approvals have been obtained.
Yes
The details of the IRB/oversight body that provided approval or exemption for the research described are given below:
The study used only openly available data in the following repositories - Kaggle, Mendeley
I confirm that all necessary patient/participant consent has been obtained and the appropriate institutional forms have been archived, and that any patient/participant/sample identifiers included were not known to anyone (e.g., hospital staff, patients or participants themselves) outside the research group so cannot be used to identify individuals.
Yes
I understand that all clinical trials and any other prospective interventional studies must be registered with an ICMJE-approved registry, such as ClinicalTrials.gov. I confirm that any such study reported in the manuscript has been registered and the trial registration ID is provided (note: if posting a prospective study registered retrospectively, please provide a statement in the trial ID field explaining why the study was not registered in advance).
Yes
I have followed all appropriate research reporting guidelines, such as any relevant EQUATOR Network research reporting checklist(s) and other pertinent material, if applicable.
Yes
Data Availability
BUSC_Mendeley - https://data.mendeley.com/datasets/wmy84gzngw/1 -Rodrigues, Paulo Sergio (2017), Breast Ultrasound Image, Mendeley Data, V1, doi: 10.17632/wmy84gzngw.1 BUSI_corrected - https://www.kaggle.com/datasets/jarintasnim090/busi-corrected - Al-Dhabyani W, Gomaa M, Khaled H, Fahmy A. Dataset of breast ultrasound images. Data in Brief. 2020 Feb;28:104863. DOI: 10.1016/j.dib.2019.104863. RODTOOk - origin - http://www.onlinemedicalimages.com/index.php/en/site-map -Rodtook, A., Kirimasthong, K., Lohitvisate, W., Makhanov, S.S. (2018) Automatic initialization of active contours and level set method in ultrasound images of breast abnormalities. Pattern Recognition, Vol 79, pp 172-182". QAMEBI - origin - https://qamebi.com/breast-ultrasound-images-database/ - [1] A. Abbasian Ardakani, A. Mohammadi, M. Mirza-Aghazadeh-Attari, U.R. Acharya, An open-access breast lesion ultrasound image database: Applicable in artificial intelligence studies, Computers in Biology and Medicine, 152 (2023) 106438. https://doi.org/10.1016/j.compbiomed.2022.106438 - [2] H. Hamyoon, W. Yee Chan, A. Mohammadi, T. Yusuf Kuzan, M. Mirza-Aghazadeh-Attari, W.L. Leong, K. Murzoglu Altintoprak, A. Vijayananthan, K. Rahmat, N. Ab Mumin, S. Sam Leong, S. Ejtehadifar, F. Faeghi, J. Abolghasemi, E.J. Ciaccio, U. Rajendra Acharya, A. Abbasian Ardakani, Artificial intelligence, BI-RADS evaluation and morphometry: A novel combination to diagnose breast cancer using ultrasonography, results from multi-center cohorts, European Journal of Radiology, 157 (2022) 110591. https://doi.org/10.1016/j.ejrad.2022.110591