PT - JOURNAL ARTICLE AU - MacCarthy, Gideon AU - Pazoki, Raha TI - Using machine learning to evaluate the value of genetic liabilities in classification of hypertension within the UK Biobank AID - 10.1101/2024.03.18.24304461 DP - 2024 Jan 01 TA - medRxiv PG - 2024.03.18.24304461 4099 - http://medrxiv.org/content/early/2024/03/18/2024.03.18.24304461.short 4100 - http://medrxiv.org/content/early/2024/03/18/2024.03.18.24304461.full AB - Background and objective Hypertension increases the risk of cardiovascular diseases (CVD) such as stroke, heart attack, heart failure, and kidney disease, contributing to global disease burden and premature mortality. Previous studies have utilized statistical and machine learning techniques to develop hypertension prediction models. Only a few have included genetic liabilities and evaluated their predictive values. This study aimed to develop an effective hypertension prediction model and investigate the potential influence of genetic liability for risk factors linked to CVD on hypertension risk using Random Forest (RF) and Neural Network (NN).Materials and methods The study included 244,718 participants of European ancestry. Genetic liabilities were constructed using previously identified genetic variants associated with various cardiovascular risk factors through genome-wide association studies (GWAS). The sample was randomly split into training and testing sets at a 70:30 ratio. We used RF and NN techniques to develop prediction models in the training set with or without feature selection. We evaluated the models’ discrimination performance using the area under the curve (AUC), calibration, and net reclassification improvement in the testing set.Results The models without genetic liabilities achieved AUCs of 0.70 and 0.72 using RF and NN methods, respectively. Adding genetic liabilities resulted in a modest improvement in the AUC for RF but not for NN. The best prediction model was achieved using RF (AUC =0.71, Spiegelhalter z score= 0.10, P-value= 0.92, calibration slope=0.99) constructed in stage two.Conclusion Incorporating genetic factors in the model may provide a modest incremental value for hypertension prediction beyond baseline characteristics. Our study highlighted the importance of genetic liabilities for both total cholesterol and LDL within the same prediction model adds value to the classification of hypertension.Competing Interest StatementThe authors have declared no competing interest.Funding StatementThis study did not receive any fundingAuthor DeclarationsI confirm all relevant ethical guidelines have been followed, and any necessary IRB and/or ethics committee approvals have been obtained.YesThe details of the IRB/oversight body that provided approval or exemption for the research described are given below:Division of Biomedical Sciences, Department of Life Sciences, College of Health and Life Sciences, Brunel University LondonI confirm that all necessary patient/participant consent has been obtained and the appropriate institutional forms have been archived, and that any patient/participant/sample identifiers included were not known to anyone (e.g., hospital staff, patients or participants themselves) outside the research group so cannot be used to identify individuals.YesI understand that all clinical trials and any other prospective interventional studies must be registered with an ICMJE-approved registry, such as ClinicalTrials.gov. I confirm that any such study reported in the manuscript has been registered and the trial registration ID is provided (note: if posting a prospective study registered retrospectively, please provide a statement in the trial ID field explaining why the study was not registered in advance).YesI have followed all appropriate research reporting guidelines, such as any relevant EQUATOR Network research reporting checklist(s) and other pertinent material, if applicable.YesAll data produced in the present study are available upon reasonable request to the authors