Skip to main content
medRxiv
  • Home
  • About
  • Submit
  • ALERTS / RSS
Advanced Search

A prediction model based on machine learning for diagnosing the early COVID-19 patients

Nan-Nan Sun, Ya Yang, Ling-Ling Tang, Yi-Ning Dai, Hai-Nv Gao, Hong-Ying Pan, Bin Ju
doi: https://doi.org/10.1101/2020.06.03.20120881
Nan-Nan Sun
1Hangzhou Wowjoy Information Technology Co., Ltd, Hangzhou, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Ya Yang
3State Key Laboratory for Diagnosis and Treatment of Infectious Diseases, National Clinical Research Centre for Infectious Diseases, Collaborative Innovation Centre for Diagnosis and Treatment of Infectious Diseases, the First Affiliated Hospital, College of Medicine, Zhejiang University, Hangzhou, Zhejiang Province, 310003, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Ling-Ling Tang
4Department of Infectious Diseases, ShuLan (Hangzhou) Hospital Affiliated to Zhejiang Shuren University Shulan International Medical College, Hangzhou, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Yi-Ning Dai
2Department of Infectious Diseases, Zhejiang Provincial People’s Hospital, People’s Hospital of Hangzhou Medical College, Hangzhou, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Hai-Nv Gao
4Department of Infectious Diseases, ShuLan (Hangzhou) Hospital Affiliated to Zhejiang Shuren University Shulan International Medical College, Hangzhou, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Hong-Ying Pan
2Department of Infectious Diseases, Zhejiang Provincial People’s Hospital, People’s Hospital of Hangzhou Medical College, Hangzhou, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
Bin Ju
1Hangzhou Wowjoy Information Technology Co., Ltd, Hangzhou, China
  • Find this author on Google Scholar
  • Find this author on PubMed
  • Search for this author on this site
  • Abstract
  • Full Text
  • Info/History
  • Metrics
  • Data/Code
  • Preview PDF
Loading

Abstract

With the dramatically fast spread of COVID-9, real-time reverse transcription polymerase chain reaction (RT-PCR) test has become the gold standard method for confirmation of COVID-19 infection. However, RT-PCR tests are complicated in operation andIt usually takes 5-6 hours or even longer to get the result. Additionally, due to the low virus loads in early COVID-19 patients, RT-PCR tests display false negative results in a number of cases. Analyzing complex medical datasets based on machine learning provides health care workers excellent opportunities for developing a simple and efficient COVID-19 diagnostic system. This paper aims at extracting risk factors from clinical data of early COVID-19 infected patients and utilizing four types of traditional machine learning approaches including logistic regression(LR), support vector machine(SVM), decision tree(DT), random forest(RF) and a deep learning-based method for diagnosis of early COVID-19. The results show that the LR predictive model presents a higher specificity rate of 0.95, an area under the receiver operating curve (AUC) of 0.971 and an improved sensitivity rate of 0.82, which makes it optimal for the screening of early COVID-19 infection. We also perform the verification for generality of the best model (LR predictive model) among Zhejiang population, and analyze the contribution of the factors to the predictive models. Our manuscript describes and highlights the ability of machine learning methods for improving the accuracy and timeliness of early COVID-19 infection diagnosis. The higher AUC of our LR-base predictive model makes it a more conducive method for assisting COVID-19 diagnosis. The optimal model has been encapsulated as a mobile application (APP) and implemented in some hospitals in Zhejiang Province.

Competing Interest Statement

The authors have declared no competing interest.

Funding Statement

2020C03123-2 the emergency project of key research and development plan in Zhejiang Province 2017ZX10204401 National Science and Technology Major 253 Project

Author Declarations

I confirm all relevant ethical guidelines have been followed, and any necessary IRB and/or ethics committee approvals have been obtained.

Yes

The details of the IRB/oversight body that provided approval or exemption for the research described are given below:

the Ethics Committee of Zhejiang Provincial People's Hospital

All necessary patient/participant consent has been obtained and the appropriate institutional forms have been archived.

Yes

I understand that all clinical trials and any other prospective interventional studies must be registered with an ICMJE-approved registry, such as ClinicalTrials.gov. I confirm that any such study reported in the manuscript has been registered and the trial registration ID is provided (note: if posting a prospective study registered retrospectively, please provide a statement in the trial ID field explaining why the study was not registered in advance).

Yes

I have followed all appropriate research reporting guidelines and uploaded the relevant EQUATOR Network research reporting checklist(s) and other pertinent material as supplementary files, if applicable.

Yes

Data Availability

all data are fully available without restriction

Copyright 
The copyright holder for this preprint is the author/funder, who has granted medRxiv a license to display the preprint in perpetuity. All rights reserved. No reuse allowed without permission.
Back to top
PreviousNext
Posted June 04, 2020.
Download PDF
Data/Code
Email

Thank you for your interest in spreading the word about medRxiv.

NOTE: Your email address is requested solely to identify you as the sender of this article.

Enter multiple addresses on separate lines or separate them with commas.
A prediction model based on machine learning for diagnosing the early COVID-19 patients
(Your Name) has forwarded a page to you from medRxiv
(Your Name) thought you would like to see this page from the medRxiv website.
CAPTCHA
This question is for testing whether or not you are a human visitor and to prevent automated spam submissions.
Share
A prediction model based on machine learning for diagnosing the early COVID-19 patients
Nan-Nan Sun, Ya Yang, Ling-Ling Tang, Yi-Ning Dai, Hai-Nv Gao, Hong-Ying Pan, Bin Ju
medRxiv 2020.06.03.20120881; doi: https://doi.org/10.1101/2020.06.03.20120881
Digg logo Reddit logo Twitter logo CiteULike logo Facebook logo Google logo Mendeley logo
Citation Tools
A prediction model based on machine learning for diagnosing the early COVID-19 patients
Nan-Nan Sun, Ya Yang, Ling-Ling Tang, Yi-Ning Dai, Hai-Nv Gao, Hong-Ying Pan, Bin Ju
medRxiv 2020.06.03.20120881; doi: https://doi.org/10.1101/2020.06.03.20120881

Citation Manager Formats

  • BibTeX
  • Bookends
  • EasyBib
  • EndNote (tagged)
  • EndNote 8 (xml)
  • Medlars
  • Mendeley
  • Papers
  • RefWorks Tagged
  • Ref Manager
  • RIS
  • Zotero
  • Tweet Widget
  • Facebook Like
  • Google Plus One

Subject Area

  • Infectious Diseases (except HIV/AIDS)
Subject Areas
All Articles
  • Addiction Medicine (70)
  • Allergy and Immunology (169)
  • Anesthesia (51)
  • Cardiovascular Medicine (456)
  • Dentistry and Oral Medicine (83)
  • Dermatology (55)
  • Emergency Medicine (160)
  • Endocrinology (including Diabetes Mellitus and Metabolic Disease) (192)
  • Epidemiology (5307)
  • Forensic Medicine (3)
  • Gastroenterology (198)
  • Genetic and Genomic Medicine (762)
  • Geriatric Medicine (81)
  • Health Economics (214)
  • Health Informatics (703)
  • Health Policy (363)
  • Health Systems and Quality Improvement (224)
  • Hematology (100)
  • HIV/AIDS (165)
  • Infectious Diseases (except HIV/AIDS) (5953)
  • Intensive Care and Critical Care Medicine (368)
  • Medical Education (105)
  • Medical Ethics (25)
  • Nephrology (83)
  • Neurology (777)
  • Nursing (43)
  • Nutrition (135)
  • Obstetrics and Gynecology (146)
  • Occupational and Environmental Health (236)
  • Oncology (482)
  • Ophthalmology (155)
  • Orthopedics (40)
  • Otolaryngology (98)
  • Pain Medicine (39)
  • Palliative Medicine (20)
  • Pathology (141)
  • Pediatrics (223)
  • Pharmacology and Therapeutics (138)
  • Primary Care Research (99)
  • Psychiatry and Clinical Psychology (868)
  • Public and Global Health (2043)
  • Radiology and Imaging (356)
  • Rehabilitation Medicine and Physical Therapy (160)
  • Respiratory Medicine (287)
  • Rheumatology (95)
  • Sexual and Reproductive Health (74)
  • Sports Medicine (77)
  • Surgery (110)
  • Toxicology (25)
  • Transplantation (30)
  • Urology (40)