<?xml version="1.0" encoding="UTF-8"?><!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "http://dtd.nlm.nih.gov/publishing/3.0/journalpublishing3.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" dtd-version="3.0" xml:lang="en" article-type="research article">
 <front>
  <journal-meta>
   <journal-id journal-id-type="publisher-id">
    abb
   </journal-id>
   <journal-title-group>
    <journal-title>
     Advances in Bioscience and Biotechnology
    </journal-title>
   </journal-title-group>
   <issn pub-type="epub">
    2156-8456
   </issn>
   <issn publication-format="print">
    2156-8502
   </issn>
   <publisher>
    <publisher-name>
     Scientific Research Publishing
    </publisher-name>
   </publisher>
  </journal-meta>
  <article-meta>
   <article-id pub-id-type="doi">
    10.4236/abb.2025.169025
   </article-id>
   <article-id pub-id-type="publisher-id">
    abb-145650
   </article-id>
   <article-categories>
    <subj-group subj-group-type="heading">
     <subject>
      Articles
     </subject>
    </subj-group>
    <subj-group subj-group-type="Discipline-v2">
     <subject>
      Biomedical 
     </subject>
     <subject>
       Life Sciences
     </subject>
    </subj-group>
   </article-categories>
   <title-group>
    Accurate Classification of Diabetes via PM Generative AI
   </title-group>
   <contrib-group>
    <contrib contrib-type="author" xlink:type="simple">
     <name name-style="western">
      <surname>
       Philip de
      </surname>
      <given-names>
       Melo
      </given-names>
     </name>
    </contrib>
    <contrib contrib-type="author" xlink:type="simple">
     <name name-style="western">
      <surname>
       Marie St.
      </surname>
      <given-names>
       Rose
      </given-names>
     </name>
    </contrib>
   </contrib-group> 
   <aff id="affnull">
    <addr-line>
     aDepartment of Nursing and Allied Health, Norfolk State University, Norfolk, VA, USA
    </addr-line> 
   </aff> 
   <pub-date pub-type="epub">
    <day>
     03
    </day> 
    <month>
     09
    </month>
    <year>
     2025
    </year>
   </pub-date> 
   <volume>
    16
   </volume> 
   <issue>
    09
   </issue>
   <fpage>
    379
   </fpage>
   <lpage>
    409
   </lpage>
   <history>
    <date date-type="received">
     <day>
      5,
     </day>
     <month>
      March
     </month>
     <year>
      2025
     </year>
    </date>
    <date date-type="published">
     <day>
      14,
     </day>
     <month>
      March
     </month>
     <year>
      2025
     </year> 
    </date> 
    <date date-type="accepted">
     <day>
      14,
     </day>
     <month>
      September
     </month>
     <year>
      2025
     </year> 
    </date>
   </history>
   <permissions>
    <copyright-statement>
     © Copyright 2014 by authors and Scientific Research Publishing Inc. 
    </copyright-statement>
    <copyright-year>
     2014
    </copyright-year>
    <license>
     <license-p>
      This work is licensed under the Creative Commons Attribution International License (CC BY). http://creativecommons.org/licenses/by/4.0/
     </license-p>
    </license>
   </permissions>
   <abstract>
    The recent surge in demand for timely and accurate health information has highlighted the need for more advanced data analysis tools. To reduce the incidence of preventable medical errors, sophisticated IT-driven classification and prediction algorithms are essential. However, extracting meaningful insights from complex biomedical data remains a significant challenge in healthcare transformation. Modern biomedical and health research generates diverse data types, including electronic health records (EHRs), medical imaging, sensor data, and telemedicine inputs, which are often complex, heterogeneous, poorly annotated, and largely unstructured. Traditional statistical learning and data mining methods require extensive preprocessing before developing predictive or clustering models. This process becomes even more challenging when dealing with intricate datasets and limited domain-specific knowledge. Recent advancements in deep learning offer promising end-to-end models capable of handling such complexity. However, these models do not consistently achieve the high levels of accuracy required by healthcare professionals. In this study, we introduce a novel Deep Learning Algorithm combined with a generative AI designed to improve classification accuracy in clinical applications significantly. The algorithm is tailored for seamless integration into hospital workflows and electronic health record systems—an area that is the central focus of our ongoing research. The proposed method combines real-world clinical data with synthetic data generated by Principal Model Generative AI. This approach increased classification accuracy in our experiments from 76% to 95% - 98%.
   </abstract>
   <kwd-group> 
    <kwd>
     Diabetes
    </kwd> 
    <kwd>
      Informatics
    </kwd> 
    <kwd>
      PIMA Data Set
    </kwd> 
    <kwd>
      Deep Learning
    </kwd> 
    <kwd>
      Improved Accuracy
    </kwd> 
    <kwd>
      PM Generative AI
    </kwd>
   </kwd-group>
  </article-meta>
 </front>
 <body>
  <sec id="s1">
   <title>1. Introduction</title>
   <p>According to the International Diabetes Federation (IDF), diabetes is a widespread chronic condition affecting over 380 million people globally, projected to exceed 600 million in the coming decades. Notably, many of these cases are preventable. Poor blood glucose regulation in diabetic individuals increases the risk of complications such as neuropathy, contributing to higher morbidity and mortality rates. Diabetes also remains a major cause of death due to its strong association with coronary artery disease and stroke.</p>
   <p>In 2013, global spending on diabetes care was estimated at a minimum of $550 billion, with projections indicating a rise to over $630 billion by 2035.</p>
   <p>Various information technology (IT)-driven solutions have been introduced to address these challenges to enhance blood glucose monitoring and diabetes management. Research shows that IT interventions can improve metabolic control and support the comprehensive care of individuals with chronic diabetes. A literature review in <xref ref-type="bibr" rid="scirp.145650-1">
     [1]
    </xref> emphasized the potential of technology-based solutions to foster efficient and informative communication between patients and healthcare providers.</p>
   <p>In addition, developing end-to-end learning models from complex datasets is essential for advancing diabetes care. However, deep learning algorithms can sometimes lack accuracy. This paper proposes a novel hybrid approach that boosts accuracy by aggregating randomly selected data subsets, thereby significantly enhancing the performance of analytical algorithms.</p>
   <p>Technology-driven interventions offer numerous benefits in healthcare, including reduced medical errors, improved research through data generation, and enhanced capacity for continuous quality improvement. Nevertheless, these innovations also come with challenges, such as high implementation and maintenance costs, usability issues for healthcare providers, and the potential for reduced face-to-face patient interaction <xref ref-type="bibr" rid="scirp.145650-2">
     [2]
    </xref>.</p>
   <p>Recent studies suggest that IT-based interventions can lead to better glycemic control and more effective diabetes management, though their impact may vary across clinical outcomes. Future research could focus on integrating multiple IT tools into unified systems to improve clinical outcomes and overall diabetes care.</p>
   <p>Information technologies also play a crucial role in classifying and diagnosing diabetes using data analytics, machine learning, and artificial intelligence. Key technologies include <xref ref-type="bibr" rid="scirp.145650-3">
     [3]
    </xref>:</p>
   <p>1) Machine Learning &amp; Artificial Intelligence:</p>
   <p>Supervised Learning: Algorithms such as decision trees, support vector machines (SVM), random forests, and neural networks classify diabetes based on patient data.</p>
   <p>Deep Learning: Convolutional Neural Networks (CNNs) and Recurrent Neural Networks (RNNs) process large datasets to enhance diagnostic accuracy.</p>
   <p>Natural Language Processing (NLP): Extracts insights from electronic health records (EHRs) and medical literature.</p>
   <p>2) Big Data and Predictive Analytics: Classify patients based on risk factors (e.g., age, BMI, glucose levels).</p>
   <p>Detect patterns for early diagnosis using historical patient records.</p>
   <p>3) Wearable Devices and Internet of Things (IoT):</p>
   <p>Continuous glucose monitors (CGMs) collect real-time blood sugar data.</p>
   <p>Smart insulin pumps adjust insulin delivery based on AI-driven predictions.</p>
   <p>Electronic Health Records (EHRs) and Cloud Computing:</p>
   <p>Store and analyze large volumes of diabetes-related data.</p>
   <p>Cloud-based AI models support real-time diagnosis and continuous monitoring.</p>
   <p>Genetic and Biomarker Analysis IT tools analyze genetic data to classify Type 1, Type 2, and gestational diabetes. Omics technologies (genomics, proteomics) assist in precision medicine.</p>
   <p>Telemedicine and Mobile Health (mHealth) Mobile apps (e.g., Glucose Buddy, MySugr) help monitor diabetes in real-time. AI-powered chatbots provide diabetes education and lifestyle recommendations.</p>
   <p>Utilizing deep learning for diabetes classification poses multiple challenges, such as:</p>
   <p>This paper introduces an advanced deep learning approach designed to enhance accuracy, as demonstrated by its application to diabetes classification. This approach, applied to PIMA Indian data, largely increased the accuracy of classification.</p>
  </sec><sec id="s2">
   <title>2. Data Description</title>
   <p>Diabetes diagnosis is contingent upon several essential parameters that assess blood glucose levels and related health indicators. The key parameters include:</p>
   <p>1) Fasting Blood Glucose (FBG) measures blood sugar levels after a minimum fasting duration of 8 hours. The classifications are as follows: Diabetes: &gt;126 mg/dL (7.0 mmol/L), Prediabetes: 100 - 125 mg/dL (5.6 - 6.9 mmol/L), and Normal: &lt;100 mg/dL (5.6 mmol/L).</p>
   <p>2) The Oral Glucose Tolerance Test (OGTT) evaluates blood sugar levels two hours post-consumption of a 75 g glucose solution. The results are categorized as: Diabetes: &gt;200 mg/dL (11.1 mmol/L), Prediabetes: 140 - 199 mg/dL (7.8 - 11.0 mmol/L), and Normal: &lt;140 mg/dL (7.8 mmol/L).</p>
   <p>3) Hemoglobin A1c (HbA1c) reflects the average blood sugar levels over the previous 2 to 3 months, with classifications as follows: Diabetes: &gt;6.5%, Prediabetes: 5.7% - 6.4%, and Normal: &lt;5.7%.</p>
   <p>4) The Random Blood Glucose Test assesses blood sugar at any time, regardless of food intake. A diabetes diagnosis is established if the level exceeds 200 mg/dL (11.1 mmol/L) alongside symptoms such as excessive thirst, frequent urination, or unexplained weight loss.</p>
   <p>5) Insulin and C-Peptide Levels are instrumental in differentiating between Type 1 and Type 2 diabetes, with low levels of both indicating Type 1 diabetes.</p>
   <p>6) Autoimmune Markers for Type 1 Diabetes, including the presence of autoantibodies such as GAD65, IA-2, and ZnT8, can confirm a Type 1 diabetes diagnosis.</p>
   <p>7) Ketone Levels, particularly in instances of Diabetic Ketoacidosis (DKA), where elevated urine or blood indicate severe insulin deficiency, are commonly associated with Type 1 diabetes.</p>
   <p>8) Body Mass Index (BMI) and Obesity: Increased BMI and obesity are significant risk factors for Type 2 diabetes.</p>
   <p>9) Blood Pressure and Cholesterol Levels: Hypertension and dyslipidemia (characterized by high LDL and low HDL) are often observed in individuals with diabetes. This paper will utilize the diabetes data available on Kaggle.</p>
   <p>The data set is characterized by the following features depicted by <xref ref-type="table" rid="table1">
     Table 1
    </xref>:</p>
   <p>The features of the PIMA data depicted in <xref ref-type="table" rid="table1">
     Table 1
    </xref> are as follows:</p>
   <p>
    <xref ref-type="fig" rid="fig1">
     Figure 1
    </xref> shows the number of patients with diabetes and without.</p>
   <p>Diabetes triggers a variety of physiological changes in the body, notably altering skin characteristics. One key study <xref ref-type="bibr" rid="scirp.145650-4">
     [4]
    </xref> aimed to analyze skin thickness in female diabetic patients and evaluate its potential as an indicator of diabetes progression. A one-way ANOVA test was employed to assess the impact of various factors on skin thickness, using a significance threshold of α ≤ 0.05. Results demonstrated a decrease in skin thickness as diabetes advanced. While insulin levels significantly influenced skin thickness, glucose levels did not show a similar effect.</p>
   <table-wrap id="table1">
    <label>
     <xref ref-type="table" rid="table1">
      Table 1
     </xref></label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Table 1. First 20 rows of the data set. The data set consists of 768 rows (patients) and 9 columns (features outcome).</title>
    </caption>
    <table class="MsoTableGrid custom-table" border="0" cellspacing="0" cellpadding="0"> 
     <tr> 
      <td class="custom-bottom-td acenter" width="4.03%"><p style="text-align:center"></p></td> 
      <td class="custom-bottom-td acenter" width="11.33%"><p style="text-align:center">Pregnancies</p></td> 
      <td class="custom-bottom-td acenter" width="9.14%"><p style="text-align:center">Glucose</p></td> 
      <td class="custom-bottom-td acenter" width="12.27%"><p style="text-align:center">Blood</p><p style="text-align:center">Pressure</p></td> 
      <td class="custom-bottom-td acenter" width="11.76%"><p style="text-align:center">Skin</p><p style="text-align:center">Thickness</p></td> 
      <td class="custom-bottom-td acenter" width="10.29%"><p style="text-align:center">Insulin</p></td> 
      <td class="custom-bottom-td acenter" width="7.36%"><p style="text-align:center">BMI</p></td> 
      <td class="custom-bottom-td acenter" width="17.65%"><p style="text-align:center">Diabetes Pedigree</p><p style="text-align:center">Function</p></td> 
      <td class="custom-bottom-td acenter" width="7.36%"><p style="text-align:center">Age</p></td> 
      <td class="custom-bottom-td acenter" width="8.82%"><p style="text-align:center">Outcome</p></td> 
     </tr> 
     <tr> 
      <td class="custom-top-td acenter" width="4.03%"><p style="text-align:center">0</p></td> 
      <td class="custom-top-td acenter" width="11.33%"><p style="text-align:center">6</p></td> 
      <td class="custom-top-td acenter" width="9.14%"><p style="text-align:center">148</p></td> 
      <td class="custom-top-td acenter" width="12.27%"><p style="text-align:center">72</p></td> 
      <td class="custom-top-td acenter" width="11.76%"><p style="text-align:center">35</p></td> 
      <td class="custom-top-td acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="custom-top-td acenter" width="7.36%"><p style="text-align:center">33.6</p></td> 
      <td class="custom-top-td acenter" width="17.65%"><p style="text-align:center">0.627</p></td> 
      <td class="custom-top-td acenter" width="7.36%"><p style="text-align:center">50</p></td> 
      <td class="custom-top-td acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">1</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">1</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">85</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">66</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">29</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">26.6</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.351</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">31</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">2</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">8</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">183</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">64</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">23.3</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.672</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">32</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">3</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">1</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">89</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">66</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">23</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">94</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">28.1</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.167</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">21</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">4</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">137</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">40</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">35</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">168</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">43.1</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">2.288</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">33</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">5</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">5</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">116</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">74</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">25.6</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.201</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">30</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">6</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">3</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">78</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">50</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">32</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">88</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">31.0</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.248</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">26</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">7</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">10</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">115</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">35.3</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.134</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">29</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">8</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">2</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">197</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">70</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">45</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">543</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">30.5</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.158</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">53</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">9</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">8</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">125</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">96</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">0.0</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.232</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">54</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">10</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">4</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">110</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">92</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">37.6</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.191</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">30</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">11</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">10</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">168</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">74</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">38.0</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.537</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">34</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">12</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">10</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">139</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">80</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">27.1</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">1.441</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">57</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">13</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">1</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">189</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">60</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">23</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">846</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">30.1</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.398</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">59</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">14</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">5</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">166</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">72</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">19</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">175</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">25.8</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.587</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">51</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">15</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">7</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">100</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">30.0</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.484</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">32</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">16</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">118</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">84</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">47</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">230</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">45.8</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.551</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">31</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">17</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">7</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">107</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">74</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">0</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">29.6</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.254</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">31</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">18</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">1</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">103</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">30</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">38</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">83</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">43.3</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.183</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">33</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">0</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="4.03%"><p style="text-align:center">19</p></td> 
      <td class="acenter" width="11.33%"><p style="text-align:center">1</p></td> 
      <td class="acenter" width="9.14%"><p style="text-align:center">115</p></td> 
      <td class="acenter" width="12.27%"><p style="text-align:center">70</p></td> 
      <td class="acenter" width="11.76%"><p style="text-align:center">30</p></td> 
      <td class="acenter" width="10.29%"><p style="text-align:center">96</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">34.6</p></td> 
      <td class="acenter" width="17.65%"><p style="text-align:center">0.529</p></td> 
      <td class="acenter" width="7.36%"><p style="text-align:center">32</p></td> 
      <td class="acenter" width="8.82%"><p style="text-align:center">1</p></td> 
     </tr> 
    </table>
   </table-wrap>
   <fig id="fig1" position="float">
    <label>Figure 1</label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 1. The number of healthy patients (blue) and diabetes patients (orange).</title>
    </caption>
    <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId15.jpeg?20250917095022" />
   </fig>
   <p>These findings suggest that skin thickness may serve as a novel marker for monitoring diabetes progression in women. However, further research is needed to validate this hypothesis. This paper will explore these results in greater detail, with the next step involving the calculation of the correlation matrix (<xref ref-type="fig" rid="fig2">
     Figure 2
    </xref>).</p>
   <fig id="fig2" position="float">
    <label>Figure 2</label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 2. Heatmap shows the correlation between features and outcomes.</title>
    </caption>
    <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId16.jpeg?20250917095020" />
   </fig>
   <p>Machine learning (ML) plays an increasingly critical role in diabetes research, treatment, and management. By analyzing patient data—such as glucose levels, body mass index (BMI), and family history—ML models can assess an individual’s risk of developing diabetes. Techniques like decision trees, neural networks, and support vector machines enable early detection.</p>
   <p>ML-based models can also predict blood glucose trends by examining historical data, dietary habits, physical activity, and medication adherence. Continuous glucose monitoring (CGM) devices, enhanced by artificial intelligence, provide real-time alerts to help prevent episodes of hypoglycemia or hyperglycemia <xref ref-type="bibr" rid="scirp.145650-4">
     [4]
    </xref>. Personalized treatment strategies are greatly improved through ML, which considers patient-specific factors, including lifestyle and medication response.</p>
   <p>Moreover, reinforcement learning algorithms assist in optimizing insulin dosages, while deep learning methods analyze retinal images for early detection of diabetic retinopathy. AI-powered tools, such as those developed by Google’s DeepMind, support healthcare professionals in diagnosing diabetes-related ocular complications.</p>
   <p>Machine learning is crucial in predicting complications such as cardiovascular disease, neuropathy, and kidney disorders, enabling timely interventions. AI-powered applications offer personalized recommendations tailored to individual dietary habits, physical activity, and sleep patterns. Wearable devices like Fitbit and Apple Watch also provide real-time data that supports more effective diabetes management. In addition, automated insulin delivery systems use machine learning to dynamically adjust insulin doses based on glucose level fluctuations, reducing patient burden and improving blood sugar control.</p>
  </sec><sec id="s3">
   <title>3. Feature Engineering</title>
   <sec id="s3_1">
    <title>3.1. Visualization</title>
    <p>Feature engineering is a crucial step in data preprocessing for machine learning. It involves identifying and transforming the variables most significantly impacting model outcomes and decision-making. This process converts raw data into meaningful features suitable for machine learning algorithms. Feature engineering includes selecting, extracting, and transforming the most relevant features from a dataset to enhance model performance and accuracy.</p>
    <p>
     <xref ref-type="fig" rid="fig2">
      Figure 2
     </xref> presents a cross-correlation heatmap highlighting key patterns and relationships within the data. This visualization aids in improving the model’s learning capability by revealing essential interdependence. After data cleaning and labeling, machine learning teams typically conduct thorough exploratory data analysis (EDA) to assess the data quality and fitness for modeling. Visual tools such as histograms, scatter plots, box plots, line graphs, and bar charts are vital in verifying data integrity. These visualizations help scientists detect data patterns, spot anomalies, test hypotheses, and validate assumptions.</p>
    <p>The performance of machine learning models is greatly affected by the quality of the features employed in their training. Feature engineering involves a range of techniques that facilitate the generation of new features by combining or transforming existing ones. These approaches are essential.</p>
    <p>Exploratory data analysis does not require formal modeling; instead, data science teams can utilize visualizations to gain meaningful insights from the data. The data collection stage involves compiling all relevant information necessary for machine learning, which can be a labor-intensive task due to the often-fragmented nature of data across various sources, such as personal computers, data warehouses, cloud storage, applications, and devices.</p>
    <p>Connecting to these diverse data sources can present considerable challenges. Furthermore, the volume of data is increasing at an extraordinary pace, leading to large datasets that demand comprehensive analysis. Additionally, the formats and types of data can differ significantly depending on their source, complicating the integration of various data types, such as video and tabular data. These figures suggest that four major features determine the outcomes. We will consider both cases when all features are considered (<xref ref-type="fig" rid="fig3">
      Figure 3
     </xref>) and only four features (<xref ref-type="fig" rid="fig4">
      Figure 4
     </xref>).</p>
    <fig id="fig3" position="float">
     <label>Figure 3</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 3. Feature engineering analysis shows the contribution of each feature in the outcome.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId17.jpeg?20250917095025" />
    </fig>
    <fig id="fig4" position="float">
     <label>Figure 4</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 4. Feature engineering analysis shows the most contribution comes from 4 features.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId18.jpeg?20250917095025" />
    </fig>
   </sec>
   <sec id="s3_2">
    <title>3.2. Detection and Removal of Outliers</title>
    <p>Outlier detection and removal is a technique used to identify and eliminate anomalies from a dataset. The main goal of this process is to improve the accuracy of data representation, which can significantly impact model performance. The extent of this impact varies—some models are highly sensitive to outliers, while others remain largely unaffected.</p>
    <p>For instance, linear regression is particularly susceptible to outliers, making it essential to address them before training a model. Various methods can be employed to manage outliers, including:</p>
    <p>Removal: This approach involves deleting records containing outliers from the dataset. However, if outliers are present in multiple variables, this can lead to substantial data loss.</p>
    <p>Replacing Values: Here, outliers are treated as missing values and replaced with imputed values deemed appropriate.</p>
    <p>Capping: This method substitutes extreme values with a predetermined threshold, or a value based on the variable’s distribution.</p>
    <p>Discretization: This process converts continuous variables into discrete categories by segmenting the range into intervals or bins.</p>
    <p>The following boxplots illustrate the selected features. In descriptive statistics, a boxplot (also called a box-and-whisker plot) is a valuable tool for exploration data analysis. It provides a visual representation of data distribution and skewness by displaying quartiles, percentiles, and averages. A boxplot encapsulates the five-number summary of a dataset, consisting of:</p>
    <p>Minimum Score: The lowest value in the dataset, excluding outliers, represented by the left whisker’s endpoint.</p>
    <p>Lower Quartile (Q1): The first quartile, marking the value below which 25% of the data falls.</p>
    <p>Median (Q2): The midpoint of the dataset, splitting it into two equal halves. It signifies that half of the values are above and half are below this point.</p>
    <p>Upper Quartile (Q3): The third quartile, representing the value below which 75% of the data lies.</p>
    <p>Maximum Score: The highest value in the dataset, excluding outliers, shown at the right whisker’s endpoint.</p>
    <p>Whiskers: These extend from the box, covering the lower and upper 25% of the data, excluding extreme outliers.</p>
    <p>Interquartile Range (IQR): The range between the first and third quartiles, representing the middle 50% of the data.</p>
    <p>Boxplots efficiently summarize data using a simple visual format, making it easy to identify quartiles, medians, and outliers briefly. Outliers can significantly affect data analysis and machine learning models, leading to biased predictions, misleading conclusions, and distorted statistical measures.</p>
    <p>To mitigate these effects, statistical methods such as the Interquartile Range (IQR) are used to quantify dispersion. The IQR measures the spread of the middle 50% of the data and is calculated as the difference between the 75th percentile (Q3) and the 25th percentile (Q1). The following figures show the boxplots of major features and outliers. <xref ref-type="fig" rid="fig5">
      Figure 5
     </xref> shows Q1, Q2, Q3 quartiles and IQR (shadowed). Outliers are beyond the upper bound.</p>
    <fig id="fig5" position="float">
     <label>Figure 5</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 5. Outliers detected in the BMI records.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId19.jpeg?20250917095028" />
    </fig>
    <p>
     <xref ref-type="fig" rid="fig6">
      Figure 6
     </xref> shows the boxplot for the Glucose records with outlier at 0 and <xref ref-type="fig" rid="fig7">
      Figure 7
     </xref> demonstrates the Boxplot for the Pregnancy records.</p>
    <fig id="fig6" position="float">
     <label>Figure 6</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 6. Outliers detected in the Glucose records.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId20.jpeg?20250917095028" />
    </fig>
    <fig id="fig7" position="float">
     <label>Figure 7</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 7. Outliers in the Pregnancies records.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId21.jpeg?20250917095029" />
    </fig>
    <p>
     <xref ref-type="fig" rid="fig8">
      Figure 8
     </xref> presents the boxplot of the Age data. Although it indicates outliers beyond the upper bound, we will not exclude them in this case, as the cohort includes patients aged 60 - 70.</p>
    <fig id="fig8" position="float">
     <label>Figure 8</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 8. Outliers detected in the Age records, but they were included in the analysis.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId22.jpeg?20250917095028" />
    </fig>
    <p>The Interquartile Range (IQR) is a statistical measure of dispersion that represents the spread of the middle 50% of a dataset. It is calculated by subtracting the 25th percentile (Q1) from the 75th percentile (Q3). To detect outliers using the IQR method, two boundaries are established:</p>
    <p>These boundaries help identify potential outliers in a dataset. Any data point below the lower bound (Q1 – 1.5 × IQR) is considered an outlier, as it significantly deviates from the rest of the data and may warrant further review or removal. Similarly, any data point above the upper bound (Q3 + 1.5 × IQR) is considered an outlier, as it is substantially higher than most data points and may require special attention.</p>
    <p>One of the main advantages of the IQR method is its resilience to skewed data distributions. Since it identifies outliers based on percentiles, it is less sensitive to extreme values. Additionally, the IQR method is straightforward to implement and interpret, providing a clear range within which most data points should fall. This makes it an effective tool for data analysis and quality control.</p>
   </sec>
   <sec id="s3_3">
    <title>3.3. Z-Score to Drop Outliers</title>
    <p>Outliers can arise from various factors and often result from either genuine variability in the data or errors in data collection, measurement, or recording. Common causes of outliers include:</p>
    <p>The Z-score is used to standardize variables, giving insight into how far a particular observation is from the meaning. Specifically, the Z-score indicates how many standard deviations a data point is away from the mean. The process of transforming a feature into Z-scores is known as standardization.</p>
    <p>If the Z-score of a data point exceeds 3, it suggests the data point is significantly different from the others and could be an outlier. Z-scores can be both positive and negative; the farther the score is from 0, the higher the likelihood that the data point is an outlier. Generally, a Z-score greater than 3 is considered extreme.</p>
    <p>For the Z-score method to be effective in identifying outliers, the data should follow a normal distribution. Removing outliers using the Z-score method involves the following steps:</p>
    <p>
     <xref ref-type="bibr" rid="scirp.145650-"></xref> 
     <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
       <mi>
         Z 
       </mi> 
       <mo>
         = 
       </mo> 
       <mfrac> 
        <mrow> 
         <mi>
           X 
         </mi> 
         <mo>
           − 
         </mo> 
         <mi>
           μ 
         </mi> 
        </mrow> 
        <mi>
          σ 
        </mi> 
       </mfrac> 
      </mrow> 
     </math></p>
    <p>where: X = data point, µ = mean of the dataset, σ = standard deviation of the dataset The procedure includes the following steps:</p>
    <p>Choose a threshold (commonly 2 or 3)</p>
    <p>A common choice is Z &gt; 3 or Z &lt; −3, meaning the data point is 3 standard deviations away from the mean. If the dataset is small or you want to be less strict, you might use Z &gt; 2.5 or even Z &gt; 2. Remove outliers: Any data point with a Z-score beyond the chosen threshold is considered an outlier and can be removed.</p>
    <p>The Z-score indicates the number of standard deviations a data point is from the mean. It is calculated as: Z-score removal of outliers requires normal distribution.</p>
    <fig id="fig9" position="float">
     <label>Figure 9</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 9. Z-scores are useful for detecting outliers—data points that deviate significantly from the rest of the dataset. Generally, data points with z-scores exceeding 3 or falling below -3 are considered potential outliers and may require further analysis.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId25.jpeg?20250917095032" />
    </fig>
    <p>
     <xref ref-type="fig" rid="fig9">
      Figure 9
     </xref> shows the application of z-scoring to data sets which after the z-transform becomes normally distributed. While Z-scores can be computed for any distribution, they are particularly useful when the data is normally distributed, because:</p>
    <p>Z-scores don’t require a normal distribution but are most useful when the data is (approximately) normal. <xref ref-type="fig" rid="fig10">
      Figure 10
     </xref> shows the detection of outliers using z-scores. <xref ref-type="fig" rid="fig11">
      Figure 11
     </xref>, <xref ref-type="fig" rid="fig12">
      Figure 12
     </xref> show the feature distributions: <xref ref-type="fig" rid="fig11">
      Figure 11
     </xref> represents skewed normal distributions and can be used to eliminate outliers, while <xref ref-type="fig" rid="fig12">
      Figure 12
     </xref> does not represent normal distribution and cannot use a-scores to eliminate outliers.</p>
    <fig id="fig10" position="float">
     <label>Figure 10</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 10. Moderately unusual and evident outliers.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId26.jpeg?20250917095031" />
    </fig>
    <fig id="fig11" position="float">
     <label>Figure 11</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 11. The data distribution for BMI and Glucose represent slightly skewed normal distributions and can be used in z-scoring for outliers’ removal.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId27.jpeg?20250917095030" />
    </fig>
    <fig id="fig12" position="float">
     <label>Figure 12</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 12. The data distribution for Pregnancies and Age cannot be used for z-scoring in outliers removal.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId28.jpeg?20250917095030" />
    </fig>
    <p>
     <xref ref-type="fig" rid="fig13">
      Figure 13
     </xref> shows the BMI and Glucose features after outliers have been removed using z-scores. It is important to emphasize that outliers can disproportionately impact models, particularly those sensitive to extreme values (e.g., linear regression, k-means clustering). In classification tasks, outliers can lead to misclassification by distorting decision boundaries. A model trained on data with outliers may struggle to generalize effectively to unseen data. Tree-based methods, such as Random Forest and XGBoost, exhibit reduced sensitivity to outliers. However, in the context of machine learning applied to public health or healthcare, it is crucial to remove outliers to avoid degrading model accuracy.</p>
    <fig id="fig13" position="float">
     <label>Figure 13</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 13. The data distribution for BMI and Glucose after elimination of outliers.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId29.jpeg?20250917095030" />
    </fig>
   </sec>
  </sec><sec id="s4">
   <title>
    <xref ref-type="bibr" rid="scirp.145650-"></xref>4. Diabetes Features against Outcomes</title>
   <sec id="s4_1">
    <title>4.1. BMI vs. Diabetes</title>
    <p>Having obesity significantly increases the likelihood of developing diabetes, a condition characterized by excessive glucose (sugar) in the bloodstream. It also accelerates the progression of diabetes. A high BMI (in the overweight or obese range) raises the risk of insulin resistance, which can lead to type 2 diabetes. Excess fat, particularly around the abdomen, disrupts insulin function, making it harder for cells to absorb glucose from the blood <xref ref-type="bibr" rid="scirp.145650-5">
      [5]
     </xref>.</p>
    <p>Research indicates that individuals with a BMI of 30 or higher face a much greater risk of developing type 2 diabetes compared to those with a normal BMI (18.5 - 24.9). Those in the overweight range (BMI 25 - 29.9) are more susceptible to prediabetes, a state where blood sugar levels are elevated but not yet high enough for a diabetes diagnosis. Losing just 5% - 10% of body weight can significantly lower this risk.</p>
    <p>Here’s how it works: The pancreas regulates blood glucose levels by producing insulin, a hormone responsible for moving glucose out of the bloodstream. Under normal conditions, insulin helps transport glucose to muscles for immediate energy use or stores it in the liver for future needs.</p>
    <p>However, in cases of diabesity (a combination of obesity and diabetes), cells become resistant to insulin, preventing glucose from entering. Additionally, the liver’s glucose storage becomes saturated with fat. With nowhere else to go, glucose remains in the bloodstream, prompting the pancreas to produce more insulin to counteract this resistance.</p>
    <p>Over time, the pancreas becomes overworked and starts producing less insulin, leading to diabetes, which can quickly worsen if insulin resistance persists.</p>
    <p>Individuals with obesity are approximately six times more likely to develop type 2 diabetes than those at a healthy weight. However, not everyone with obesity will necessarily develop diabetes. Other contributing factors include:</p>
    <p>Some people with obesity may produce enough insulin without overburdening the pancreas, while others may have a limited insulin production capacity, making them more susceptible to diabesity.</p>
    <p>
     <xref ref-type="fig" rid="fig14">
      Figure 14
     </xref> demonstrates the cross plot between BMI values of patients with diabetes (pink color) and healthy ones (green color).</p>
    <fig id="fig14" position="float">
     <label>Figure 14</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 14. BMI near 30 threatens the development of diabetes.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId30.jpeg?20250917095035" />
    </fig>
   </sec>
   <sec id="s4_2">
    <title>4.2. Glucose vs. Diabetes</title>
    <p>The term “glucose” comes from the Greek word meaning “sweet.” It is a type of sugar derived from the foods you consume, serving as a primary energy source for your body. When glucose circulates in your bloodstream to reach your cells, it is known as blood sugar. Insulin, a hormone, helps transfer glucose into cells for energy and storage. In the case of diabetes, blood glucose levels become higher than normal. This can happen either because your body doesn’t produce enough insulin to regulate glucose or because your cells don’t respond effectively to insulin.</p>
    <p>Prolonged high blood sugar can damage the heart and blood vessels, increasing the risk of heart disease, high blood pressure, and stroke. It can also lead to kidney disease or failure (diabetic nephropathy), damage to the retina (diabetic retinopathy), causing vision loss or blindness, and nerve damage (diabetic neuropathy), which can result in pain, tingling, or numbness, especially in the feet. Additionally, high blood sugar raises the risk of cognitive decline and Alzheimer’s disease, contributes to fatty liver disease, and worsens insulin resistance.</p>
    <p>The glucose in your bloodstream primarily comes from carbohydrate-rich foods, such as bread, potatoes, and fruit. As shown in <xref ref-type="fig" rid="fig15">
      Figure 15
     </xref>, when glucose levels exceed 125 - 130 mg/dL, the risk of developing diabetes significantly increases.</p>
    <fig id="fig15" position="float">
     <label>Figure 15</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 15. Glucose vs. diabetes shows numbers with elevated risks of diabetes.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId31.jpeg?20250917095038" />
    </fig>
    <p>As we eat, food travels down the esophagus into the stomach, where acids and enzymes break it down into smaller components, releasing glucose in the process. This glucose then moves to the intestines, where it is absorbed into the bloodstream. Insulin plays a crucial role in helping glucose enter cells. After eating, blood sugar levels naturally rise and gradually decrease a few hours later as insulin facilitates glucose absorption.</p>
    <p>Hyperglycemia, or high blood glucose, occurs when blood sugar levels exceed 200 mg/dL two hours after eating or 125 mg/dL while fasting. For patients with diabetes, regular blood sugar testing is essential. Exercise, a balanced diet, and medication can help maintain healthy blood glucose levels and prevent complications. <xref ref-type="fig" rid="fig16">
      Figure 16
     </xref> presents a violin plot showing that higher glucose levels are associated with diabetes.</p>
    <fig id="fig16" position="float">
     <label>Figure 16</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 16. Violin plot shows the glucose values vs. the outcomes.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId32.jpeg?20250917095038" />
    </fig>
    <p>Chronic high blood glucose can lead to nerve damage (diabetic neuropathy), causing pain, tingling, or numbness, especially in the feet. It also increases the risk of cognitive decline and Alzheimer’s disease, contributes to fatty liver disease, and worsens insulin resistance. The glucose in the bloodstream primarily comes from carbohydrate-rich foods, such as bread, potatoes, and fruit. <xref ref-type="fig" rid="fig15">
      Figure 15
     </xref> illustrates that when glucose levels exceed 125 - 130 mg/dL, the risk of diabetes increases significantly.</p>
   </sec>
   <sec id="s4_3">
    <title>4.3. Age vs. Diabetes</title>
    <p>Diabetes is mostly diagnosed in individuals over the age of 45. The risk increases with age due to factors such as reduced insulin sensitivity, weight gain, and decreased physical activity. However, an increasing number of younger people are being diagnosed with diabetes, largely due to lifestyle changes. More than 90% of individuals with diabetes have Type 2.</p>
    <p>For middle-aged individuals, it’s essential to focus on weight management, blood sugar control, and regular monitoring for complications.</p>
    <p>For older adults, the risk of complications like heart disease, nerve damage, and kidney problems rises, necessitating careful management and monitoring.</p>
    <p>Diabetes can go undiagnosed for years, as symptoms like excessive thirst, blurred vision, and tingling in the hands and feet may develop gradually and go unnoticed.</p>
    <p>Middle age marks a significant rise in diabetes diagnoses. Approximately 15% of middle-aged Americans are diagnosed with Type 2 diabetes, which is nearly five times the rate of those under 45. Incidence increases even more as individuals age. Nearly 25% of older adults in the U.S. have been diagnosed with Type 2 diabetes, with undiagnosed cases possibly accounting for an additional 5%. This means that over one in four senior Americans live with Type 2 diabetes.</p>
    <p>Age is a significant risk factor for Type 2 diabetes. The older a patient is, the more likely they are to develop the condition. This is also true for preteens and teenagers, as the rates of diabetes in this age group have increased sharply in recent years. Type 2 diabetes is a disease caused by a combination of genetic factors and lifestyle choices. <xref ref-type="fig" rid="fig17">
      Figure 17
     </xref> displays a cross-plot of age and outcomes, with diabetes shown in red and non-diabetes in green.</p>
    <fig id="fig17" position="float">
     <label>Figure 17</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 17. Age over 30 causes a higher risk of diabetes.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId33.jpeg?20250917095041" />
    </fig>
    <p>Being overweight, having high blood pressure, and leading a sedentary lifestyle all increase the risk of Type 2 diabetes. Managing diabetes at different ages involves different approaches:</p>
    <p>Younger individuals: Focus on diet, exercise, and insulin management (for Type 1).</p>
    <p>Middle-aged individuals: Emphasize weight management, blood sugar control, and monitoring for complications.</p>
    <p>Older adults: The risk of complications such as heart disease, nerve damage, and kidney issues increase, requiring careful management.</p>
    <p>Diabetes can go undiagnosed for years. Symptoms such as excessive thirst, blurry vision, and tingling in the hands and feet may develop gradually and be overlooked.</p>
    <p>Middle age is when the number of diabetes diagnoses starts to rise significantly. Approximately 15% of middle-aged Americans are diagnosed with Type 2 diabetes, nearly five times the rate among those under 45. The rate increases even further as individuals enter their senior years. Nearly 25% of older adults in the U.S. have been diagnosed with Type 2, with undiagnosed cases possibly accounting for an additional 5%. This means that more than one in four senior Americans lives with Type 2 diabetes <xref ref-type="bibr" rid="scirp.145650-6">
      [6]
     </xref>.</p>
    <p>The disease also is affecting ever more teens and even children. Researchers believe childhood obesity and lack of exercise are among the reasons behind that trend.</p>
   </sec>
   <sec id="s4_4">
    <title>4.4. Pregnancies vs. Diabetes</title>
    <p>Pregnancy can have a significant impact on diabetes, and diabetes can also affect pregnancy. There are two main scenarios to consider:</p>
    <p>Gestational Diabetes Mellitus (GDM)—This type of diabetes develops during pregnancy, usually in the second or third trimester, and often resolves after childbirth. It occurs when the body cannot produce enough insulin to meet the increased needs of pregnancy.</p>
    <p>Pre-existing Diabetes (Type 1 or Type 2)—Women who have diabetes before pregnancy need to manage their blood sugar levels carefully to prevent complications for both them and the baby.</p>
    <p>Gestational diabetes is diabetes that a woman can develop during pregnancy. When you have diabetes, your body cannot use the sugars and starches (carbohydrates) it takes in as food to make energy. As a result, your body collects extra sugar in your blood.</p>
    <p>We don’t know all the causes of gestational diabetes. Some with gestational diabetes are overweight before getting pregnant or have diabetes in the family. From 1 in 50 to 1 in 20 pregnant women has gestational diabetes. It is more common in Native American, Alaskan Native, Hispanic, Asian, and Black women, but it is found in white women, too. <xref ref-type="fig" rid="fig18">
      Figure 18
     </xref> shows the Pregnancies data distribution for patients with diabetes (red) and without (green).</p>
    <fig id="fig18" position="float">
     <label>Figure 18</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 18. Higher number of pregnancies indicate risks of diabetes. the KDE plot for Pregnancies showing non-zero values for values &lt; 0 is a common side effect of kernel density estimation as it smooths the data using Gaussian kernels).</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId34.jpeg?20250917095043" />
    </fig>
   </sec>
  </sec><sec id="s5">
   <title>
    <xref ref-type="bibr" rid="scirp.145650-"></xref>5. Deep Learning Analysis of Diabetes Data</title>
   <p>Classification is the process of finding or discovering a model or function that helps separate the data into multiple categorical classes, i.e., discrete values. In classification, data is categorized under different labels according to some parameters given in the input, and then the labels are predicted for the data.</p>
   <p>In a classification task, we predict discrete class labels using independent features. In the classification task, we should find a decision boundary that can separate the different classes in the target variable. For more information on advanced regression/classification algorithms, we send the reader to <xref ref-type="bibr" rid="scirp.145650-7">
     [7]
    </xref>. The derived mapping function could be demonstrated as “IF-THEN” rules. The classification process deals with problems where the data can be divided into binary or multiple discrete labels. For example, suppose we want to predict the possibility of winning a match by Team A based on some parameters recorded earlier. Then, there would be two labels: Yes and No.</p>
   <p>Regression is finding a model or function for distinguishing the data into continuous real values instead of using classes or discrete values. It can also identify the distribution movement that depends on historical data. Because a regression predictive model predicts a quantity, the skill of the model must be reported as an error in those predictions.</p>
   <p>Deep learning, or hierarchical learning, is a subset of machine learning in AI that mimics the brain’s computing abilities and decision-making patterns. In contrast to task-based algorithms, deep learning systems learn from data representations. It can learn from unstructured or unlabeled data. A neural network with multiple hidden layers and nodes in each hidden layer is known as a deep learning system or a deep neural network. i.e. depth of the neural network. Essentially, every neural network with more than three layers, that is, including the Input Layer and Output Layer can be considered a Deep Learning Model.</p>
   <p>Deep learning is the field of artificial intelligence (AI) that teaches computers to process data in a way inspired by the human brain. Deep learning models can recognize data patterns like complex pictures, text, and sounds to produce accurate insights and predictions. Neural networks power deep learning. It consists of interconnected nodes or neurons in a layered structure. The nodes process data in a coordinated and adaptive system.</p>
   <p>Neural networks form the foundation of deep learning systems. They exchange feedback on generated outputs, learn from mistakes, and improve continuously. While deep learning models are powerful, simpler neural networks are often favored for basic machine learning (ML) tasks due to their lower development costs and modest computational requirements. These simpler models are particularly feasible for smaller projects, enabling organizations to develop internal applications for tasks such as data visualization and pattern recognition cost-effectively <xref ref-type="bibr" rid="scirp.145650-8">
     [8]
    </xref>.</p>
   <p>In contrast, deep learning systems have a broad range of practical applications. Their capacity to learn from data, extract complex patterns, and automatically develop features enables them to deliver state-of-the-art performance. Common use cases include natural language processing (NLP), autonomous driving, and speech recognition <xref ref-type="bibr" rid="scirp.145650-9">
     [9]
    </xref>.</p>
   <p>However, training and developing deep learning systems require substantial computational resources and funding. As a result, many organizations opt to use pre-trained models offered as fully managed services, which can be customized for specific applications.</p>
   <p>To better understand neural networks, consider them as a series of algorithms inspired by the structure and function of the human brain. These networks are designed to recognize patterns in data, interpreting sensory inputs by labeling and clustering raw information. Their ability to learn and adapt over time makes them essential for image and speech recognition tasks.</p>
   <p>This learning path provides a comprehensive introduction to neural networks, preparing you to explore more advanced applications.</p>
   <p>
    <xref ref-type="fig" rid="fig19">
     Figure 19
    </xref> represents a simple neural network and a deep learning neural network architecture with an activation function.</p>
   <fig id="fig19" position="float">
    <label>Figure 19</label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 19. A simple neural network and a deep learning neural network characterized by deep forward and backpropagation.</title>
    </caption>
    <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId35.jpeg?20250917095045" />
   </fig>
   <p>Neural networks learn and identify patterns directly from data without relying on predefined rules. They consist of several key components:</p>
   <p>Learning Rule: The method used to adjust weights and biases over time to improve performance.</p>
   <p>An L-layer neural network can be defined as a nested function:</p>
   <p>
    <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
      <msup> 
       <mi>
         a 
       </mi> 
       <mrow> 
        <mrow> 
         <mo>
           ( 
         </mo> 
         <mi>
           L 
         </mi> 
         <mo>
           ) 
         </mo> 
        </mrow> 
       </mrow> 
      </msup> 
      <mo>
        = 
      </mo> 
      <msup> 
       <mi>
         f 
       </mi> 
       <mrow> 
        <mrow> 
         <mo>
           ( 
         </mo> 
         <mi>
           L 
         </mi> 
         <mo>
           ) 
         </mo> 
        </mrow> 
       </mrow> 
      </msup> 
      <mrow> 
       <mo>
         ( 
       </mo> 
       <mrow> 
        <msup> 
         <mi>
           W 
         </mi> 
         <mrow> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mi>
             L 
           </mi> 
           <mo>
             ) 
           </mo> 
          </mrow> 
         </mrow> 
        </msup> 
        <mi>
          f 
        </mi> 
        <mrow> 
         <mo>
           ( 
         </mo> 
         <mrow> 
          <msup> 
           <mi>
             W 
           </mi> 
           <mrow> 
            <mrow> 
             <mo>
               ( 
             </mo> 
             <mrow> 
              <mi>
                L 
              </mi> 
              <mo>
                − 
              </mo> 
              <mn>
                1 
              </mn> 
             </mrow> 
             <mo>
               ) 
             </mo> 
            </mrow> 
           </mrow> 
          </msup> 
          <mi>
            f 
          </mi> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mrow> 
            <mo>
              ⋯ 
            </mo> 
            <mi>
              f 
            </mi> 
            <mrow> 
             <mo>
               ( 
             </mo> 
             <mrow> 
              <msup> 
               <mi>
                 W 
               </mi> 
               <mrow> 
                <mrow> 
                 <mo>
                   ( 
                 </mo> 
                 <mn>
                   1 
                 </mn> 
                 <mo>
                   ) 
                 </mo> 
                </mrow> 
               </mrow> 
              </msup> 
              <mi>
                x 
              </mi> 
              <mo>
                + 
              </mo> 
              <msup> 
               <mi>
                 b 
               </mi> 
               <mrow> 
                <mrow> 
                 <mo>
                   ( 
                 </mo> 
                 <mn>
                   1 
                 </mn> 
                 <mo>
                   ) 
                 </mo> 
                </mrow> 
               </mrow> 
              </msup> 
             </mrow> 
             <mo>
               ) 
             </mo> 
            </mrow> 
            <mo>
              ⋯ 
            </mo> 
           </mrow> 
           <mo>
             ) 
           </mo> 
          </mrow> 
          <mo>
            + 
          </mo> 
          <msup> 
           <mi>
             b 
           </mi> 
           <mrow> 
            <mrow> 
             <mo>
               ( 
             </mo> 
             <mrow> 
              <mi>
                L 
              </mi> 
              <mo>
                − 
              </mo> 
              <mn>
                1 
              </mn> 
             </mrow> 
             <mo>
               ) 
             </mo> 
            </mrow> 
           </mrow> 
          </msup> 
         </mrow> 
         <mo>
           ) 
         </mo> 
        </mrow> 
        <mo>
          + 
        </mo> 
        <msup> 
         <mi>
           b 
         </mi> 
         <mrow> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mi>
             L 
           </mi> 
           <mo>
             ) 
           </mo> 
          </mrow> 
         </mrow> 
        </msup> 
       </mrow> 
       <mo>
         ) 
       </mo> 
      </mrow> 
     </mrow> 
    </math> (1)</p>
   <p>where we used 
    <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
      <msup> 
       <mi>
         f 
       </mi> 
       <mrow> 
        <mrow> 
         <mo>
           ( 
         </mo> 
         <mi>
           L 
         </mi> 
         <mo>
           ) 
         </mo> 
        </mrow> 
       </mrow> 
      </msup> 
     </mrow> 
    </math> to allow for possibly different activation functions in different layers. This nested expression explicitly shows the feedforward computation as a nested function of x applying each layer in sequence.</p>
   <p>We minimize the loss function for regression:</p>
   <p>
    <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
      <mi>
        L 
      </mi> 
      <mo>
        = 
      </mo> 
      <mfrac> 
       <mn>
         1 
       </mn> 
       <mi>
         N 
       </mi> 
      </mfrac> 
      <mstyle displaystyle="true"> 
       <msubsup> 
        <mo>
          ∑ 
        </mo> 
        <mrow> 
         <mi>
           i 
         </mi> 
         <mo>
           = 
         </mo> 
         <mn>
           1 
         </mn> 
        </mrow> 
        <mi>
          N 
        </mi> 
       </msubsup> 
       <mrow> 
        <msup> 
         <mrow> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mrow> 
            <msub> 
             <mover accent="true"> 
              <mi>
                y 
              </mi> 
              <mo>
                ^ 
              </mo> 
             </mover> 
             <mi>
               i 
             </mi> 
            </msub> 
            <mo>
              − 
            </mo> 
            <msub> 
             <mi>
               y 
             </mi> 
             <mi>
               i 
             </mi> 
            </msub> 
           </mrow> 
           <mo>
             ) 
           </mo> 
          </mrow> 
         </mrow> 
         <mn>
           2 
         </mn> 
        </msup> 
       </mrow> 
      </mstyle> 
     </mrow> 
    </math> (2)</p>
   <p>and classification problems:</p>
   <p>
    <math display="inline" xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
      <mi>
        L 
      </mi> 
      <mo>
        = 
      </mo> 
      <mo>
        − 
      </mo> 
      <mfrac> 
       <mn>
         1 
       </mn> 
       <mi>
         N 
       </mi> 
      </mfrac> 
      <mstyle displaystyle="true"> 
       <msubsup> 
        <mo>
          ∑ 
        </mo> 
        <mrow> 
         <mi>
           i 
         </mi> 
         <mo>
           = 
         </mo> 
         <mn>
           1 
         </mn> 
        </mrow> 
        <mi>
          N 
        </mi> 
       </msubsup> 
       <mrow> 
        <mrow> 
         <mo>
           [ 
         </mo> 
         <mrow> 
          <msub> 
           <mi>
             y 
           </mi> 
           <mi>
             i 
           </mi> 
          </msub> 
          <mi>
            log 
          </mi> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mrow> 
            <msub> 
             <mover accent="true"> 
              <mi>
                y 
              </mi> 
              <mo>
                ^ 
              </mo> 
             </mover> 
             <mi>
               i 
             </mi> 
            </msub> 
           </mrow> 
           <mo>
             ) 
           </mo> 
          </mrow> 
          <mo>
            + 
          </mo> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mrow> 
            <mn>
              1 
            </mn> 
            <mo>
              − 
            </mo> 
            <msub> 
             <mi>
               y 
             </mi> 
             <mi>
               i 
             </mi> 
            </msub> 
           </mrow> 
           <mo>
             ) 
           </mo> 
          </mrow> 
          <mi>
            log 
          </mi> 
          <mrow> 
           <mo>
             ( 
           </mo> 
           <mrow> 
            <mn>
              1 
            </mn> 
            <mo>
              − 
            </mo> 
            <msub> 
             <mover accent="true"> 
              <mi>
                y 
              </mi> 
              <mo>
                ^ 
              </mo> 
             </mover> 
             <mi>
               i 
             </mi> 
            </msub> 
           </mrow> 
           <mo>
             ) 
           </mo> 
          </mrow> 
         </mrow> 
         <mo>
           ] 
         </mo> 
        </mrow> 
       </mrow> 
      </mstyle> 
     </mrow> 
    </math> (3)</p>
   <p>Weights and biases are updated using the following expressions:</p>
   <p>
    <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
      <msup> 
       <mi>
         W 
       </mi> 
       <mi>
         l 
       </mi> 
      </msup> 
      <mo>
        : 
      </mo> 
      <mo>
        = 
      </mo> 
      <msup> 
       <mi>
         W 
       </mi> 
       <mi>
         l 
       </mi> 
      </msup> 
      <mo>
        − 
      </mo> 
      <mi>
        α 
      </mi> 
      <mfrac> 
       <mrow> 
        <mo>
          ∂ 
        </mo> 
        <mi>
          L 
        </mi> 
       </mrow> 
       <mrow> 
        <mo>
          ∂ 
        </mo> 
        <msup> 
         <mi>
           W 
         </mi> 
         <mi>
           l 
         </mi> 
        </msup> 
       </mrow> 
      </mfrac> 
     </mrow> 
    </math> (4)</p>
   <p>
    <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
      <msup> 
       <mi>
         b 
       </mi> 
       <mi>
         l 
       </mi> 
      </msup> 
      <mo>
        : 
      </mo> 
      <mo>
        = 
      </mo> 
      <msup> 
       <mi>
         b 
       </mi> 
       <mi>
         l 
       </mi> 
      </msup> 
      <mo>
        − 
      </mo> 
      <mi>
        α 
      </mi> 
      <mfrac> 
       <mrow> 
        <mo>
          ∂ 
        </mo> 
        <mi>
          L 
        </mi> 
       </mrow> 
       <mrow> 
        <mo>
          ∂ 
        </mo> 
        <msup> 
         <mi>
           b 
         </mi> 
         <mi>
           l 
         </mi> 
        </msup> 
       </mrow> 
      </mfrac> 
     </mrow> 
    </math> (5)</p>
   <p>The whole procedure consists of the following steps:</p>
   <p>1) Compute the loss function</p>
   <p>2) Perform forward propagation to compute activations.</p>
   <p>3) Compute gradients using backpropagation.</p>
   <p>4) Update weights and biases using gradient descent.</p>
   <p>If we want to solve regression or classification problem, the last layer of a neural network usually contains only one unit. If the activation function of the last unit is linear then the neural network is a regression model if the activation function is a logistic function the neural network is a binary classification model. The results of the diabetes data (accuracy) are shown in <xref ref-type="fig" rid="fig20">
     Figure 20
    </xref>.</p>
   <fig id="fig20" position="float">
    <label>Figure 20</label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 20. Confusion matrix for conventional deep learning.</title>
    </caption>
    <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId48.jpeg?20250917095046" />
   </fig>
   <p>
    <xref ref-type="table" rid="table2">
     Table 2
    </xref> depicts the Classification report of a conventional deep-learning algorithm applied to the diabetes data. The accuracy of the algorithm is 0.7597402597402597 which is quite typical for classification approaches (<xref ref-type="fig" rid="fig21">
     Figure 21
    </xref>).</p>
   <table-wrap id="table2">
    <label>
     <xref ref-type="table" rid="table2">
      Table 2
     </xref></label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Table 2. Classification report for a conventional deep learning algorithm. The accuracy is 76%.</title>
    </caption>
    <table class="MsoTableGrid custom-table" border="0" cellspacing="0" cellpadding="0"> 
     <tr> 
      <td class="custom-bottom-td acenter" width="20.10%"><p style="text-align:center">Outcome</p></td> 
      <td class="custom-bottom-td acenter" width="23.00%"><p style="text-align:center">Precision</p></td> 
      <td class="custom-bottom-td acenter" width="21.56%"><p style="text-align:center">Recall</p></td> 
      <td class="custom-bottom-td acenter" width="17.24%"><p style="text-align:center">f1-score</p></td> 
      <td class="custom-bottom-td acenter" width="18.09%"><p style="text-align:center">Support</p></td> 
     </tr> 
     <tr> 
      <td class="custom-top-td acenter" width="20.10%"><p style="text-align:center">0 (negative)</p></td> 
      <td class="custom-top-td acenter" width="23.00%"><p style="text-align:center">0.80</p></td> 
      <td class="custom-top-td acenter" width="21.56%"><p style="text-align:center">0.83</p></td> 
      <td class="custom-top-td acenter" width="17.24%"><p style="text-align:center">0.82</p></td> 
      <td class="custom-top-td acenter" width="18.09%"><p style="text-align:center">99</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="20.10%"><p style="text-align:center">1 (Positive)</p></td> 
      <td class="acenter" width="23.00%"><p style="text-align:center">0.67</p></td> 
      <td class="acenter" width="21.56%"><p style="text-align:center">0.64</p></td> 
      <td class="acenter" width="17.24%"><p style="text-align:center">0.65</p></td> 
      <td class="acenter" width="18.09%"><p style="text-align:center">55</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="20.10%"><p style="text-align:center">Accuracy</p></td> 
      <td class="acenter" width="61.81%" colspan="3"><p style="text-align:center">76%</p></td> 
      <td class="acenter" width="18.09%"><p style="text-align:center">154</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="20.10%"><p style="text-align:center">Macro avg</p></td> 
      <td class="acenter" width="23.00%"><p style="text-align:center">0.74</p></td> 
      <td class="acenter" width="21.56%"><p style="text-align:center">0.73</p></td> 
      <td class="acenter" width="17.24%"><p style="text-align:center">0.74</p></td> 
      <td class="acenter" width="18.09%"><p style="text-align:center">154</p></td> 
     </tr> 
     <tr> 
      <td class="acenter" width="20.10%"><p style="text-align:center">Weighted avg</p></td> 
      <td class="acenter" width="23.00%"><p style="text-align:center">0.76</p></td> 
      <td class="acenter" width="21.56%"><p style="text-align:center">0.76</p></td> 
      <td class="acenter" width="17.24%"><p style="text-align:center">0.76</p></td> 
      <td class="acenter" width="18.09%"><p style="text-align:center">154</p></td> 
     </tr> 
    </table>
   </table-wrap>
   <fig id="fig21" position="float">
    <label>Figure 21</label>
    <caption>
     <title>
      <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 21. Accuracy of other methods applied to the data.</title>
    </caption>
    <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId49.jpeg?20250917095044" />
   </fig>
   <p>The confusion matrix is given by <xref ref-type="fig" rid="fig20">
     Figure 20
    </xref>:</p>
   <p>Metric Description:</p>
   <p>Precision: Proportion of positive identifications that were correct;</p>
   <p>Recall: Proportion of actual positives that were correctly identified;</p>
   <p>F1-score: Harmonic mean of precision and recall;</p>
   <p>Support: Number of actual occurrences of the class in the dataset.</p>
   <p>Deep learning has various challenges:</p>
   <p>Data Challenges</p>
   <p>Computational Challenges</p>
   <p>Let us apply other methods to the diabetes data. In supervised learning, the model is trained on labeled data, meaning the input data comes with corresponding output labels. The goal is to learn a mapping from input to outputs.</p>
   <p>We will apply the following supervised learning algorithms:</p>
   <p>1) Logistic Regression (for classification tasks);</p>
   <p>2) Decision Trees;</p>
   <p>3) Random Forests;</p>
   <p>4) Support Vector Machines (SVM);</p>
   <p>5) KNeighborsClassifier.</p>
   <p>KNN is a simple, supervised machine learning (ML) algorithm that can be used for classification or regression tasks - and is also frequently used in missing value imputation. It is based on the idea that the observations closest to a given data point are the most “similar” observations in a data set, and we can therefore classify unforeseen points based on the values of the closest existing points. By choosing K, the user can select the number of nearby observations to use in the algorithm <xref ref-type="bibr" rid="scirp.145650-10">
     [10]
    </xref>. The accuracy of all known methods is in the interval (72% - 26%) (<xref ref-type="fig" rid="fig21">
     Figure 21
    </xref>).</p>
   <p>In the next section, we will consider an optimized deep learning algorithm that improves the accuracy to up to 95% and higher with the same architecture and training conditions.</p>
  </sec><sec id="s6">
   <title>
    <xref ref-type="bibr" rid="scirp.145650-"></xref>6. Generative Artificial Intelligence Algorithms</title>
   <sec id="s6_1">
    <title>6.1. Clinical Significance of Generative AIs</title>
    <p>The literature shows the effectiveness of generative AI in clinical decision-making, as well as highlighting its challenges. For example, AI systems provided an effective method for the management of gastrointestinal disease including early detection and diagnosis, and the author suggested using various methods of generative AI for different gastrointestinal health information in future research <xref ref-type="bibr" rid="scirp.145650-11">
      [11]
     </xref>. Furthermore, AI has already demonstrated great capability in breast cancer management and patient outcomes and . For example, AI systems have shown promise in mammography screening and radiotherapy treatment . In addition to the benefits of AI, the researchers noted some of its challenges such as patient confidentiality, problems with moral principles, and regulation issues . They emphasized the need to conduct specific studies that can prove the accuracy of AI techniques in healthcare delivery .</p>
    <p>Despite these challenges, AI has the potential of making a dramatic change in the delivery of healthcare services by empowering health practitioners to provide the best patient care <xref ref-type="bibr" rid="scirp.145650-14">
      [14]
     </xref>. Similarly, AI should be held in high regard because of its ability to identify not only specific diseases but all diseases and healthcare professionals should pay attention to its effectiveness <xref ref-type="bibr" rid="scirp.145650-15">
      [15]
     </xref>. Other studies highlighted the potential of Generative AI into clinical practice, specifically in the prevention of cardiovascular disease <xref ref-type="bibr" rid="scirp.145650-16">
      [16]
     </xref>.</p>
    <p>Moreover, AI has proven to be effective in clinical decision-making despite limitations. Similarly, AI techniques have the potential of accurately classifying diabetes to improve the health status of patients.</p>
   </sec>
   <sec id="s6_2">
    <title>6.2. Principal Model Generative Artificial Intelligence Algorithm (PM GenAI)</title>
    <p>In this paper, we developed a new generative AI algorithm that consists of the following steps: 1) PM GenAI explores the data and generates new data by sampling from the posterior distribution of the model weights. This allows it to create diverse outputs based on the learned distributions, rather than relying on a fixed, deterministic model, 2) Uncertainty Quantification: By modeling uncertainty in parameters, PM GenAI produces a distribution of outputs instead of a single prediction. This is particularly useful for tasks that benefit from multiple plausible outcomes, such as data generation or decision-making under uncertainty, 3) Design the neural network architecture and specify prior distributions for the model parameters. Use techniques such as variational inference or Markov Chain Monte Carlo (MCMC) to approximate the posterior distribution of the weights, 4) Once the posterior is approximated (typically via backpropagation), sample weights from the distribution and perform feedforward passes through the network to generate predictions or synthetic data points that reflect the learned uncertainty.</p>
    <p>The algorithm divides the data into test and training. In the training data set it defines the trend and statistics: Mean vector (μ), the center of the distribution, covariance matrix (Σ): the shape, spread, and orientation. The PDF of a Gaussian component is:</p>
    <p>
     <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
       <mi>
         P 
       </mi> 
       <mrow> 
        <mo>
          ( 
        </mo> 
        <mrow> 
         <mrow> 
          <mi>
            x 
          </mi> 
          <mo>
            | 
          </mo> 
         </mrow> 
         <msub> 
          <mi>
            μ 
          </mi> 
          <mi>
            k 
          </mi> 
         </msub> 
         <mo>
           , 
         </mo> 
         <msub> 
          <mi>
            Σ 
          </mi> 
          <mi>
            k 
          </mi> 
         </msub> 
        </mrow> 
        <mo>
          ) 
        </mo> 
       </mrow> 
       <mo>
         = 
       </mo> 
       <mfrac> 
        <mn>
          1 
        </mn> 
        <mrow> 
         <msup> 
          <mrow> 
           <mrow> 
            <mo>
              ( 
            </mo> 
            <mrow> 
             <mn>
               2 
             </mn> 
             <mi>
               π 
             </mi> 
            </mrow> 
            <mo>
              ) 
            </mo> 
           </mrow> 
          </mrow> 
          <mrow> 
           <mrow> 
            <mn>
              1 
            </mn> 
            <mo>
              / 
            </mo> 
            <mi>
              d 
            </mi> 
           </mrow> 
          </mrow> 
         </msup> 
         <msup> 
          <mrow> 
           <mrow> 
            <mo>
              | 
            </mo> 
            <mrow> 
             <msub> 
              <mi>
                Σ 
              </mi> 
              <mi>
                k 
              </mi> 
             </msub> 
            </mrow> 
            <mo>
              | 
            </mo> 
           </mrow> 
          </mrow> 
          <mrow> 
           <mrow> 
            <mn>
              1 
            </mn> 
            <mo>
              / 
            </mo> 
            <mn>
              2 
            </mn> 
           </mrow> 
          </mrow> 
         </msup> 
        </mrow> 
       </mfrac> 
       <mi>
         exp 
       </mi> 
       <mrow> 
        <mo>
          ( 
        </mo> 
        <mrow> 
         <mo>
           − 
         </mo> 
         <mfrac> 
          <mn>
            1 
          </mn> 
          <mn>
            2 
          </mn> 
         </mfrac> 
         <msup> 
          <mrow> 
           <mrow> 
            <mo>
              ( 
            </mo> 
            <mrow> 
             <mi>
               x 
             </mi> 
             <mo>
               − 
             </mo> 
             <msub> 
              <mi>
                μ 
              </mi> 
              <mi>
                k 
              </mi> 
             </msub> 
            </mrow> 
            <mo>
              ) 
            </mo> 
           </mrow> 
          </mrow> 
          <mtext>
            T 
          </mtext> 
         </msup> 
         <msubsup> 
          <mi>
            Σ 
          </mi> 
          <mi>
            k 
          </mi> 
          <mrow> 
           <mo>
             − 
           </mo> 
           <mn>
             1 
           </mn> 
          </mrow> 
         </msubsup> 
         <mrow> 
          <mo>
            ( 
          </mo> 
          <mrow> 
           <mi>
             x 
           </mi> 
           <mo>
             − 
           </mo> 
           <msub> 
            <mi>
              μ 
            </mi> 
            <mi>
              k 
            </mi> 
           </msub> 
          </mrow> 
          <mo>
            ) 
          </mo> 
         </mrow> 
        </mrow> 
        <mo>
          ) 
        </mo> 
       </mrow> 
      </mrow> 
     </math> (6)</p>
    <p>The covariance matrix is:</p>
    <p>
     <math xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
       <msub> 
        <mi>
          Σ 
        </mi> 
        <mi>
          k 
        </mi> 
       </msub> 
       <mo>
         = 
       </mo> 
       <mfrac> 
        <mn>
          1 
        </mn> 
        <mrow> 
         <msub> 
          <mi>
            N 
          </mi> 
          <mi>
            k 
          </mi> 
         </msub> 
        </mrow> 
       </mfrac> 
       <mstyle displaystyle="true"> 
        <msubsup> 
         <mo>
           ∑ 
         </mo> 
         <mrow> 
          <mi>
            i 
          </mi> 
          <mo>
            = 
          </mo> 
          <mn>
            1 
          </mn> 
         </mrow> 
         <mi>
           N 
         </mi> 
        </msubsup> 
        <mrow> 
         <msub> 
          <mi>
            γ 
          </mi> 
          <mrow> 
           <mi>
             i 
           </mi> 
           <mi>
             k 
           </mi> 
          </mrow> 
         </msub> 
         <mrow> 
          <mo>
            ( 
          </mo> 
          <mrow> 
           <msub> 
            <mi>
              x 
            </mi> 
            <mi>
              i 
            </mi> 
           </msub> 
           <mo>
             − 
           </mo> 
           <msub> 
            <mi>
              μ 
            </mi> 
            <mi>
              k 
            </mi> 
           </msub> 
          </mrow> 
          <mo>
            ) 
          </mo> 
         </mrow> 
         <msup> 
          <mrow> 
           <mrow> 
            <mo>
              ( 
            </mo> 
            <mrow> 
             <msub> 
              <mi>
                x 
              </mi> 
              <mi>
                i 
              </mi> 
             </msub> 
             <mo>
               − 
             </mo> 
             <msub> 
              <mi>
                μ 
              </mi> 
              <mi>
                k 
              </mi> 
             </msub> 
            </mrow> 
            <mo>
              ) 
            </mo> 
           </mrow> 
          </mrow> 
          <mtext>
            T 
          </mtext> 
         </msup> 
        </mrow> 
       </mstyle> 
      </mrow> 
     </math> (7)</p>
    <p>In this expression:</p>
    <p>
     <math display="inline" xmlns="http://www.w3.org/1998/Math/MathML"> <mrow> 
       <msub> 
        <mi>
          N 
        </mi> 
        <mi>
          k 
        </mi> 
       </msub> 
       <mo>
         = 
       </mo> 
       <mstyle displaystyle="true"> 
        <msubsup> 
         <mo>
           ∑ 
         </mo> 
         <mrow> 
          <mi>
            i 
          </mi> 
          <mo>
            = 
          </mo> 
          <mn>
            1 
          </mn> 
         </mrow> 
         <mi>
           N 
         </mi> 
        </msubsup> 
        <mrow> 
         <msub> 
          <mi>
            γ 
          </mi> 
          <mrow> 
           <mi>
             i 
           </mi> 
           <mi>
             κ 
           </mi> 
          </mrow> 
         </msub> 
        </mrow> 
       </mstyle> 
      </mrow> 
     </math> (8)</p>
    <p>PM GenAI generates augmented data that resembles real data, which is especially valuable in scenarios with limited labeled samples largely improving the quality of data sets. By incorporating uncertainty, the algorithm also enables models to be more robust to outliers and noise, ultimately improving generalization. In summary, using PM GenAI in deep learning for data generation enables the incorporation of uncertainty into models, generate realistic synthetic data, and make more reliable and accurate predictions.</p>
   </sec>
   <sec id="s6_3">
    <title>6.3. The Choice of Hyperparameters</title>
    <p>Learning Rate: Controls how much the model updates its weight during training. A high learning rate may lead to convergence issues, while a low one can make training slow (we use Adam or Adagrad optimizers to automatically adjust the learning rate).</p>
    <p>Batch Size: Defines the number of training samples used in one forward and backward pass. Common choices include 32, 64, and 128. We use 128 and then split it into mini batches.</p>
    <p>Number of Epochs: Specifies how many times the entire dataset is passed through the neural network. We usually do not use more than 100 epochs.</p>
    <p>Optimizer (e.g., SGD, Adam, RMSprop): Determines how the model updates weights during training. We usually use an Adam optimizer.</p>
    <p>Loss Function (e.g., MSE, Cross-Entropy): Measures the difference between predicted and actual outputs. We use Cross-Entropy because we solve the classification problem. MSE us used in regression algorithms,</p>
    <p>Number of Layers and Neurons per Layer: Defines the depth and complexity of the network. We start with 8 neurons (the number of features). It is highly advisable this way of building the neural network to prevent overfitting if many neurons are selected.</p>
    <p>Dropout Rate: Controls the percentage of neurons randomly dropped during training to prevent overfitting (we use dropout rate 0.3-0.4 to prevent over or underfitting)</p>
    <p>Weight Initialization: Determines how weights are initialized before training (e.g., Xavier, He initialization). Choosing the right weight initialization method depends on the activation function and network depth. We use Xavier initialization to work well with classification problems. Proper weight initialization leads to more stable training and improved model performance.</p>
    <p>Activation Functions (e.g., ReLU, Sigmoid): Define how neurons process inputs. We use ReLU for hidden layers and Sigmoid for the output layer typical to classification problems.</p>
    <p>L1/L2 Regularization (Weight Decay): Helps prevent overfitting by penalizing large weights. It is implemented in optimizers like AdamW and SGD.</p>
    <p>K-fold cross-validation. We used K-fold cross-validation as a resampling technique to evaluate the performance of the method. It evaluates how well a model generalizes unseen data by splitting the dataset into multiple subsets. K-fold cross-validation includes:</p>
    <p>1) The dataset is divided into K equally sized folds (subsets).</p>
    <p>2) The model is trained on K-1 folds and tested on the remaining one.</p>
    <p>3) The process is repeated K times, with each fold serving as the test set once.</p>
    <p>4) The final model performance is obtained by averaging the results across all K iterations.</p>
    <p>5) Stratified K-Fold ensures class distribution is preserved in each fold (useful for imbalanced datasets and can be used in several applications including public health).</p>
    <p>Statistical analysis of the results shows that the comparison of previous methods (<xref ref-type="fig" rid="fig21">
      Figure 21
     </xref>) and the PM GenAI accuracy yields:</p>
    <p>P ~ 0.004 &lt; 0.05</p>
    <p>We may conclude that PM GenAI significantly improves the performance of the classification algorithms for diabetes.</p>
   </sec>
   <sec id="s6_4">
    <title>6.4. PM GenAI Confusion Matrix and Accuracy</title>
    <p>
     <xref ref-type="fig" rid="fig22">
      Figure 22
     </xref> depicts the Confusion matrix generated by PM GenAI. <xref ref-type="table" rid="table3">
      Table 3
     </xref> is a complete classification report demonstrating 97% accuracy.</p>
    <fig id="fig22" position="float">
     <label>Figure 22</label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Figure 22. Confusion matrix shows large improvement of accuracy compared with other methods.</title>
     </caption>
     <graphic mimetype="image" position="float" xlink:type="simple" xlink:href="https://html.scirp.org/file/7302221-rId56.jpeg?20250917095057" />
    </fig>
    <table-wrap id="table3">
     <label>
      <xref ref-type="table" rid="table3">
       Table 3
      </xref></label>
     <caption>
      <title>
       <xref ref-type="bibr" rid="scirp.145650-"></xref>Table 3. Classification report for a PM GenAI algorithm. The accuracy is 97%.</title>
     </caption>
     <table class="MsoTableGrid custom-table" border="0" cellspacing="0" cellpadding="0"> 
      <tr> 
       <td class="custom-bottom-td acenter" width="20.10%"><p style="text-align:center">Outcome</p></td> 
       <td class="custom-bottom-td acenter" width="20.09%"><p style="text-align:center">Precision</p></td> 
       <td class="custom-bottom-td acenter" width="19.93%"><p style="text-align:center">Recall</p></td> 
       <td class="custom-bottom-td acenter" width="19.87%"><p style="text-align:center">f1-score</p></td> 
       <td class="custom-bottom-td acenter" width="20.01%"><p style="text-align:center">Support</p></td> 
      </tr> 
      <tr> 
       <td class="custom-top-td acenter" width="20.10%"><p style="text-align:center">0 (negative)</p></td> 
       <td class="custom-top-td acenter" width="20.09%"><p style="text-align:center">0.95</p></td> 
       <td class="custom-top-td acenter" width="19.93%"><p style="text-align:center">0.98</p></td> 
       <td class="custom-top-td acenter" width="19.87%"><p style="text-align:center">0.97</p></td> 
       <td class="custom-top-td acenter" width="20.01%"><p style="text-align:center">1949</p></td> 
      </tr> 
      <tr> 
       <td class="acenter" width="20.10%"><p style="text-align:center">1 (Positive)</p></td> 
       <td class="acenter" width="20.09%"><p style="text-align:center">0.98</p></td> 
       <td class="acenter" width="19.93%"><p style="text-align:center">0.95</p></td> 
       <td class="acenter" width="19.87%"><p style="text-align:center">0.97</p></td> 
       <td class="acenter" width="20.01%"><p style="text-align:center">1826</p></td> 
      </tr> 
      <tr> 
       <td class="acenter" width="20.10%"><p style="text-align:center">Accuracy</p></td> 
       <td class="acenter" width="59.89%" colspan="3"><p style="text-align:center">97%</p></td> 
       <td class="acenter" width="20.01%"><p style="text-align:center">3775</p></td> 
      </tr> 
      <tr> 
       <td class="acenter" width="20.10%"><p style="text-align:center">Macro avg</p></td> 
       <td class="acenter" width="20.09%"><p style="text-align:center">0.97</p></td> 
       <td class="acenter" width="19.93%"><p style="text-align:center">0.96</p></td> 
       <td class="acenter" width="19.87%"><p style="text-align:center">0.97</p></td> 
       <td class="acenter" width="20.01%"><p style="text-align:center">3775</p></td> 
      </tr> 
      <tr> 
       <td class="acenter" width="20.10%"><p style="text-align:center">Weighted avg</p></td> 
       <td class="acenter" width="20.09%"><p style="text-align:center">0.97</p></td> 
       <td class="acenter" width="19.93%"><p style="text-align:center">0.97</p></td> 
       <td class="acenter" width="19.87%"><p style="text-align:center">0.97</p></td> 
       <td class="acenter" width="20.01%"><p style="text-align:center">3775</p></td> 
      </tr> 
     </table>
    </table-wrap>
   </sec>
  </sec><sec id="s7">
   <title>
    <xref ref-type="bibr" rid="scirp.145650-"></xref>7. Conclusions</title>
   <p>The abundance of biomedical data presents significant opportunities and challenges in healthcare research. A crucial aspect is identifying relationships among diverse data points to develop reliable medical tools using data-driven techniques and machine learning. Previous studies have integrated multiple data sources to achieve this and create comprehensive knowledge bases for predictive analysis and discovery. While existing models show considerable potential, machine learning-based predictive tools have yet to gain widespread adoption in the medical field.</p>
   <p>Research has demonstrated that optimized algorithms, such as Support Vector Machines (SVM), play a vital role in accurately diagnosing various diseases <xref ref-type="bibr" rid="scirp.145650-17">
     [17]
    </xref>.</p>
   <p>Despite these advancements, deep learning approaches remain underutilized in addressing a wide range of healthcare and medical challenges despite their significant potential. Deep learning offers several advantages, including superior performance, an end-to-end learning framework with built-in feature extraction, and the ability to process complex, multi-modal data. However, for these methods to be effectively implemented in healthcare, researchers must address challenges associated with medical data, which is often sparse, noisy, heterogeneous, and time dependent <xref ref-type="bibr" rid="scirp.145650-18">
     [18]
    </xref>.</p>
   <p>Additionally, improved methodologies and tools are needed to facilitate the seamless integration of deep learning into healthcare workflows and clinical decision-making.</p>
   <p>In this study, we initially utilized a Kaggle dataset provided by the National Institute of Diabetes and Digestive and Kidney Diseases. This dataset aims to predict diabetes risk in patients based on specific diagnostic indicators. The data were curated with constraints, including only Pima Indian women aged 21 and above. It contains multiple predictor variables, including the number of pregnancies, body mass index (BMI), insulin levels, age, and other relevant factors, alongside the target outcome variable.</p>
   <p>The results of this study are auspicious, achieving an accuracy rate of approximately 95% - 98% (see also <xref ref-type="bibr" rid="scirp.145650-12">
     [12]
    </xref>). <xref ref-type="bibr" rid="scirp.145650-19">
     [19]
    </xref> compared different methods and found that deep learning gives the best results. <xref ref-type="bibr" rid="scirp.145650-20">
     [20]
    </xref> published a paper on projections of global mortality and burden of disease focusing on diabetes. <xref ref-type="bibr" rid="scirp.145650-21">
     [21]
    </xref> considered the data mining approach to prediction, while and <xref ref-type="bibr" rid="scirp.145650-23">
     [23]
    </xref> discussed the risks of disease development using ML approaches.</p>
   <p>The proposed method offers several advantages:</p>
   <p>1) It reduces variance and enhances predictive accuracy.</p>
   <p>2) It performs well on large datasets, effectively handling missing and noisy data.</p>
   <p>3) It prioritizes features based on their impact on decision-making, enabling analysts to focus on critical variables.</p>
   <p>4) It efficiently processes high-dimensional data, making it highly suitable for real-world decision support applications.</p>
  </sec>
 </body><back>
  <ref-list>
   <title>References</title>
   <ref id="scirp.145650-ref1">
    <label>1</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Adaji, A., Schattner, P. and Jones, K. (2008) The Use of Information Technology to Enhance Diabetes Management in Primary Care: A Literature Review. Journal of Innovation in Health Informatics, 16, 229-237. &gt;https://doi.org/10.14236/jhi.v16i3.698
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref2">
    <label>2</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Barnett, A.H., Eff, C., Leslie, R.D.G. and Pyke, D.A. (1981) Diabetes in Identical Twins. A Study of 200 Pairs. Diabetologia, 20, 87-93. &gt;https://doi.org/10.1007/bf00262007
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref3">
    <label>3</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Redondo, M.J., Jeffrey, J., Fain, P.R., Eisenbarth, G.S. and Orban, T. (2008) Concordance for Islet Autoimmunity among Monozygotic Twins. New England Journal of Medicine, 359, 2849-2850. &gt;https://doi.org/10.1056/nejmc0805398
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref4">
    <label>4</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Tol, A. and Baghbanian, A. (2012) The Introduction of Self-Management in Type 2 Diabetes Care: A Narrative Review. Journal of Education and Health Promotion, 1, 35. &gt;https://doi.org/10.4103/2277-9531.102048
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref5">
    <label>5</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Cho, N.H., Whiting, D., Guariguata, L., Montoya, P.A., Forouhi, N., Hambleton, I., et al., (2013) IDF Diabetes Atlas. 6th Edition, International Diabetes Federation.
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref6">
    <label>6</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Riazi, H., Larijani, B., Langarizadeh, M. and Shahmoradi, L. (2015) Managing Diabetes Mellitus Using Information Technology: A Systematic Review. Journal of Diabetes&amp;Metabolic Disorders, 14, Article No. 49. &gt;https://doi.org/10.1186/s40200-015-0174-x
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref7">
    <label>7</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Funjan, K.I. (2020) Skin Thickness Can Predict the Progress of Diabetes Type 2: A New Medical Hypothesis. EC Diabetes and Metabolic Research, 4, 8-12.
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref8">
    <label>8</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Helmer, J. (2024) How Age Relates to Type 2 Diabetes. Google Publication. &gt;https://www.webmd.com/diabetes/diabetes-link-age 
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref9">
    <label>9</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Liu, B., Song, L., Zhang, L., Wang, L., Wu, M., Xu, S., et al. (2020) Higher Numbers of Pregnancies Associated with an Increased Prevalence of Gestational Diabetes Mellitus: Results from the Healthy Baby Cohort Study. Journal of Epidemiology, 30, 208-212. &gt;https://doi.org/10.2188/jea.je20180245
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref10">
    <label>10</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     De Melo, P. (2024) Public Health Informatics and Technology. Library of Congress, Washington DC.
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref11">
    <label>11</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Lee, K. and Kim, E.S. (2024) Generative Artificial Intelligence in the Early Diagnosis of Gastrointestinal Disease. Applied Sciences, 14, Article 11219. &gt;https://doi.org/10.3390/app142311219
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref12">
    <label>12</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Ahn, J.S., Shin, S., Yang, S., Park, E.K., Kim, K.H., Cho, S.I., et al. (2023) Artificial Intelligence in Breast Cancer Diagnosis and Personalized Medicine. Journal of Breast Cancer, 26, 405-435. &gt;https://doi.org/10.4048/jbc.2023.26.e45
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref13">
    <label>13</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     AlSamhori, J.F., AlSamhori, A.R.F., Duncan, L.A., Qalajo, A., Alshahwan, H.F., Al-abbadi, M., et al. (2024) Artificial Intelligence for Breast Cancer: Implications for Diagnosis and Management. Journal of Medicine, Surgery, and Public Health, 3, Article 100120. &gt;https://doi.org/10.1016/j.glmedi.2024.100120
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref14">
    <label>14</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Aamir, A., Iqbal, A., Jawed, F., Ashfaque, F., Hafsa, H., Anas, Z., et al. (2024) Exploring the Current and Prospective Role of Artificial Intelligence in Disease Diagnosis. Annals of Medicine&amp;Surgery, 86, 943-949. &gt;https://doi.org/10.1097/ms9.0000000000001700
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref15">
    <label>15</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Kaur, S., Singla, J., Nkenyereye, L., Jha, S., Prashar, D., Joshi, G.P., et al. (2020) Medical Diagnostic Systems Using Artificial Intelligence (AI) Algorithms: Principles and Perspectives. IEEE Access, 8, 228049-228069. &gt;https://doi.org/10.1109/access.2020.3042273
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref16">
    <label>16</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Parsa, S., Somani, S., Dudum, R., Jain, S.S. and Rodriguez, F. (2024) Artificial Intelligence in Cardiovascular Disease Prevention: Is It Ready for Prime Time? Current Atherosclerosis Reports, 26, 263-272. &gt;https://doi.org/10.1007/s11883-024-01210-w
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref17">
    <label>17</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     De Melo, P. and Davtyan, M. (2023) High Accuracy Classification of Populations with Breast Cancer: SVM Approach. Cancer Research Journal, 11, 94-104.
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref18">
    <label>18</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     de Melo, P. (2025) Augmented and Synthetic Data in Artificial Intelligence. International Journal of Artificial Intelligence&amp;Applications, 16, 93-108. &gt;https://doi.org/10.5121/ijaia.2025.16307
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref19">
    <label>19</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Naz, H. and Ahuja, S. (2020) Deep Learning Approach for Diabetes Prediction Using PIMA Indian Dataset. Journal of Diabetes&amp;Metabolic Disorders, 19, 391-403. &gt;https://doi.org/10.1007/s40200-020-00520-5
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref20">
    <label>20</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Mathers, C.D. and Loncar, D. (2006) Projections of Global Mortality and Burden of Disease from 2002 to 2030. PLOS Medicine, 3, e442. &gt;https://doi.org/10.1371/journal.pmed.0030442
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref21">
    <label>21</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Swapna, G., Vinayakumar, R. and Soman, K.P. (2018) Diabetes Detection Using Deep Learning Algorithms. ICT Express, 4, 243-246. &gt;https://doi.org/10.1016/j.icte.2018.10.005
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref22">
    <label>22</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     Wu, H., Yang, S., Huang, Z., He, J. and Wang, X. (2018) Type 2 Diabetes Mellitus Prediction Model Based on Data Mining. Informatics in Medicine Unlocked, 10, 100-107. &gt;https://doi.org/10.1016/j.imu.2017.12.006
    </mixed-citation>
   </ref>
   <ref id="scirp.145650-ref23">
    <label>23</label>
    <mixed-citation publication-type="other" xlink:type="simple">
     The Emerging Risk Factors Collaboration, (2010) Diabetes Mellitus, Fasting Blood Glucose Concentration, and Risk of Vascular Disease: A Collaborative Meta-Analysis of 102 Prospective Studies. The Lancet, 375, 2215-2222. &gt;https://doi.org/10.1016/s0140-6736(10)60484-9
    </mixed-citation>
   </ref>
  </ref-list>
 </back>
</article>