<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Analysis of Datasets Created to Assess the Developing Gestational Diabetes Mellitus* Risk of</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Arabboev</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Shohruh</string-name>
          <email>bek.shohruh@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Begmatov</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Rikhsivoev</string-name>
          <email>mrikhsivoev@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Saidakmal</string-name>
          <email>saidakmalflash@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Saydiakbarov</institution>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Tashkent University of Information Technologies named after Muhammad al-Khwarizmi</institution>
          ,
          <addr-line>108 Amir Temur St., Tashkent, 100084</addr-line>
          ,
          <country country="UZ">Uzbekistan</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>In recent years, the healthcare field has seen a rise in the use of artificial intelligence. There is growing interest in applying artificial intelligence technology to the field of healthcare. To effectively predict disease and deploy proper artificial intelligence and machine learning algorithms, there is a need for suitable datasets. Datasets are widely used to assess the risk of developing diabetes, one of the most common diseases. Given the preceding, this paper reviews datasets created to assess the risk of developing gestational diabetes mellitus (GDM) used worldwide.</p>
      </abstract>
      <kwd-group>
        <kwd>eol&gt;Dataset</kwd>
        <kwd>gestational diabetes mellitus</kwd>
        <kwd>machine learning</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>In recent years, the number of people with diabetes has been increasing worldwide. Diabetes is one
of the most common diseases among the population [1]. There are common types of diabetes such as
type 1 [2], type 2 [3] and gestational diabetes [4]. Gestational Diabetes Mellitus (GDM) poses
significant health risks to both mothers and infants, making its early detection and effective
management crucial for maternal and fetal well-being. As the prevalence of GDM continues to rise
globally, there is a growing need for robust predictive models to identify women at risk. Key to
developing such models is the availability and analysis of high-quality datasets specifically tailored for
assessing GDM risk factors.</p>
      <p>Artificial intelligence (AI) is a rapidly developing field, and its application in treating diabetes may
revolutionize the approach to diagnosing and managing this chronic condition [5].</p>
      <p>Machine learning algorithms are used to support predictive models for the risk of developing
diabetes or its complications [6]. Digital therapy has proven to be an established intervention for
lifestyle therapy in the treatment of diabetes. Patients are increasingly empowered to self-manage
their diabetes, and both patients and healthcare professionals benefit from clinical decision support.
AI enables continuous and remote monitoring of patient symptoms and biomarkers. Technological
advances have helped optimize resource utilization in diabetes. Artificial intelligence is changing
massive processes in diabetes care, from traditional treatment strategies to creating targeted,
datadriven precision care.</p>
      <p>Our review provides a comprehensive resource for researchers, IT companies involved in
developing medical data, and technology companies specializing in the healthcare sector.</p>
      <p>The contributions of this survey are summarised as follows:
1. Research conducted on creating a dataset for GDM across geographic regions is presented.
2. We present a comprehensive review of research on creating a dataset for assessing the risk of
gestational diabetes mellitus (GDM). This review covers the algorithms or models used, datasets used
or created, and the results achieved.</p>
      <p>3. The dependence of the number of data in the dataset on the accuracy of the model is critically
analyzed.</p>
      <p>The rest of the paper follows this structure: The section “Analysis of studies on creating datasets
for gestational diabetes” reviewed up-to-date research done in the related field. The section “Results
obtained in the analyzed studies” compared the results gained in the analyzed studies. The section
“Conclusion” concludes this paper.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Analysis of studies on creating datasets for gestational diabetes</title>
      <p>This section reviews studies on creating datasets for gestational diabetes. In preparation for this
review paper, between January and February 2024, Google Scholar, PubMed, Science Direct, and IEEE
Xplore databases were searched for articles with the keywords “gestational, diabetes mellitus, dataset,
Artificial Intelligence”. 2557 articles were found as a result of the search, of which 15 papers met
specific inclusion.
retrospective observational study. The initial search identified 26,063 patients with the following
factors: postcode, height, weight, BMI at booking, ethnicity (self-reported), parity, glucose tolerance
test offer, test results (0 min and up to 120 min after 75 g glucose load), mode of delivery, estimated
total blood loss, gestational age, newborn weight, SCBU admission, length of postpartum stay, fetal
sex, and stillbirth are some other factors to consider.</p>
      <p>In [10], it is aimed at improving the diagnosis of gestational diabetes by using data collection
methods. Also, this study analyzed the performance of supervised learning algorithms such as ID3,
Naïve Bayes, C4.5, and Random Tree. The results of the experiment showed that the Random Tree
algorithm gave the best result with the highest accuracy and the lowest error rate. The dataset used in
this study is a clinical dataset collected from St. Isabella’s Hospital, Mylapore, Chennai, and the
National Institute of Diabetes, Digestive, and Kidney Diseases, which includes records of about 600
patients. In particular, all patients listed in the dataset were pregnant and over 21 years old.</p>
      <p>In [11], it is developed a machine learning-based prediction model for gestational diabetes (GDM)
in early pregnancy in Chinese women. This study used population-based data from 19,331 pregnant
women registered as pregnant up to 15 weeks of gestation between October 2010 and August 2012 in
Tianjin, China. The dataset is randomly divided into a training set (70%) and a testing set (30%). Risk
factors collected during enrollment were reviewed and used to build a predictive model on the
training data set. Machine learning, such as the extreme gradient boosting (XGBoost) method, was
used to develop the model.</p>
      <p>In [12], it is created a dataset based on data from 6822 pregnant women living in a geographic area
defined by three regional health boards in New Zealand. The prevalence of GDM was estimated using
four commonly used data sources. Coded clinical data on diabetes status were collected from regional
health boards and the Ministry of Health’s National Minimum Data Set, and plasma glucose results
were collected from laboratories serving the recruitment area and were coded according to the New
Zealand Society for the Study of Diabetes diagnostic criteria and collected via self-administered
diabetes status questionnaires.</p>
      <p>In another interesting study [13], the authors created a dataset based on data collected at the West
China Second Hospital in Chengdu, Sichuan. A total of 33,935 pregnant women were enrolled in the
EHR from 2013 to 2016 for experimental data. The GDM-related data of these samples contained 106
features of archival data, 23 features of audit data, 157 features of laboratory information system test
data, and 268 features of EHR first pages. After data cleaning, the authors used a filtering strategy to
preselect patients whose EHR data were associated with GDM as candidate samples, excluding
pregestational diabetes. Through this process, the authors obtained an accurate data set of 10,105 samples
with common clinical characteristics. This dataset contains 1649 GDM (positive) cases and 8456
nonGDM (negative) cases.</p>
      <p>A new dataset for gestational diabetes is created in [14]. Data used in this study were collected
from local hospitals in Mysuru, Karnataka, India. Medical records were obtained after anonymizing
patients to ensure confidentiality. The dataset was developed by keeping obstetrics and gynecology
consultants in feedback. The dataset contains information on 1352 pregnant women. The GDM
dataset was developed with the help of physicians by removing less significant and irrelevant features,
followed by data cleaning and transformation.</p>
      <p>
        In [15], it is created a dataset based on data from the general ward of Kurmitola General Hospital in
Bangladesh to test an ML model and predict diabetes for Bangladeshi patients. The authors addressed
the group of trainee doctors who participated in the data collection process. They conducted a brief
interview with the patients, and after their appropriate consent, the doctors agreed to provide relevant
information to the authors. In total, it took about th
        <xref ref-type="bibr" rid="ref14">ree weeks in November 2019</xref>
        to collect all relevant
information. This split dataset contains data from 181 patients and consists of 4 characteristics: patient
age, body mass index, number of pregnancies, and glucose concentration.
      </p>
      <p>In [16], the authors created a dataset based on real pregnancy test data from a hospital in Beijing
from 2008 to 2018. The dataset contains examination records of 120,396 pregnant women. In the entire
sample set, 18,400 pregnant women had gestational diabetes, accounting for approximately 15.28%,
and 7,518 pregnant women had gestational hypertension, accounting for approximately 6.24%.</p>
      <p>In [17], it is created a new dataset based on data from 2016 to 2018 used routinely collected
maternity and birth data for singleton pregnancies that ended in birth at Monash Health, Australia’s
largest public health service, at Universal Health Melbourne. Within the framework of the health care
system, three maternity hospitals served different ethnic populations.</p>
      <p>In another interesting study, a homicidal diabetes dataset was created in [18]. This dataset included
a total of 48,502 singleton pregnancies from January 2016 to June 2021 across the Monash Health
maternity network. The incidence of GDM was 21.3%. A randomly selected 80% dataset was used for
model development and 20% for validation. Performance, including calibration and discrimination
performance, was evaluated.</p>
      <p>
        In [19], it is created a dataset obtained from patients attending the Department of Obstetrics and
Fetal Medicine at the Hospital Parroquial de San Bernardo, Santiago, Chile [19]. The dataset included
data from 1,611 different
        <xref ref-type="bibr" rid="ref1">pregnant patients from 2019</xref>
        to 2022. The dataset is divided into three parts:
the training set (70%), the validation set (10%), and the testing set (20%). In this study, twelve different
ML models and their hyperparameters were optimized to achieve early and high predictive
performance of GDM. To improve the forecast results, the method of data augmentation was used in
training. Three methods were used to select the most suitable variables for GDM prediction. After
training, the models with the highest area under the receiver operating characteristic curve were
evaluated on the validation set. The models with the best results were evaluated as a measure of
generalization performance on the test set.
      </p>
      <p>In [20], data from pregnant Mexican women included in the “Cuido mi Embarazo” (CME) group
were used for development (107 cases, 469 controls), and data from the "Monica Pretelini Sáenz"
Maternal Perinatal Group was used for the investigation (32 cases, 199 controls) [20]. A 2-hour oral
glucose tolerance test (OGTT) with 75 g of glucose at 24-28 weeks of gestation was used to diagnose
GDM. A total of 114 single nucleotide polymorphisms with predictive power were selected for
evaluation. Blood samples collected during the OGTT were used for SNP analysis. The CME group
was randomly divided into a training dataset (70% of the group) and a testing dataset (30% of the
group). The training dataset is divided into 10 groups: 9 for building the predictive model and 1 for
validation.</p>
      <p>
        In [21], it is created a dataset obtained from a perinatal database for women who gave birth at
seven hospitals in four regions of South Korea, under the authority of the Catholic University of Korea
from January 2009 to
        <xref ref-type="bibr" rid="ref8">December 2020</xref>
        . Data on mothers’ demographic characteristics, body mass
index, blood pressure measurements, blood and urine laboratory tests, diagnoses recorded by doctors,
and prescribed medications were collected from the hospital database through electronic medical
cards. In this study, a machine learning algorithm was developed to predict gestational diabetes
mellitus (GDM) using retrospective data from 34,387 multicenter pregnancies in South Korea.
      </p>
      <p>Figure 2 shows the geographical locations of the countries where the organizations of the authors
of the studies analyzed in this review paper.</p>
    </sec>
    <sec id="sec-3">
      <title>3. Results obtained in the analyzed studies</title>
      <p>This section provides information on the results obtained in the articles analyzed in Section 2. The
table below shows information about the countries of the organizations of the authors in which the
GDM datasets were created, and the number of participants registered in the dataset.</p>
      <p>Data collected in the studies
0
20000
40000
60000
80000
100000
120000
140000
160000</p>
      <p>Figure 3 depicts the datasets developed in the analyzed studies. If we compare the number of
participants in the developed datasets with each other, the five with the highest quantities belong to
Ref. [8], Ref. [16], Ref. [18], Ref. [21], and Ref. [9], respectively. Ref. [15], Ref. [10], Ref. [20], Ref. [7],
and Ref. [14], were the five lowest quantities on the number of participants in the developed datasets,
according to the analyzed studies in this review paper.</p>
      <p>It can be seen from Table 2 that different types of Machine Learning models/algorithms were used
in the analyzed articles. In [7], six ML algorithms were used. In this study, six ML algorithms were
used. In this study, the number of participants in the dataset was 1012, and the result was 86.74 %
accuracy. In [8], Quantum Machine Learning algorithm was used to develop a prediction model. In
this research, the number of participants in the dataset was 149015, and the results were 69 % for both
precision and accuracy. In [10], ID3, Naïve Bayes, C4.5 and Random Tree algorithms were used for
developing a prediction model. In this work, the number of participants in the dataset was 600, and the
result was 93.8 % accuracy. In [11], two ML algorithms, LR and XGBoost, were used to develop a
prediction model. In this study, the number of participants in the dataset was 19331, and the results
were 76.9 % specificity and 75.7 % accuracy. In [13], five ML algorithms were used for developing a
prediction model. In this work, the number of participants in the dataset was 10105, and the result
90 % accuracy. In [14], three ML algorithms were used to develop a prediction model. In this study, the
number of participants in the dataset was 1352, and the result was 93 % accuracy. In [15], the
prediction model was developed using four ML algorithms: KNN, DT, RF, and NB. In this research, the
number of participants in the dataset was 181, and the results were 80 % precision, 84 % AUC and 81.2
% accuracy. A prediction model was developed in [16] using LR, XGBoost, and LightGBM algorithms.
In this study, the dataset had 120396 participants. The results showed 91.7% accuracy, 75.3% precision,
and 92.1% AUC. In [18], twelve ML algorithms were used to develop a prediction model. In this study,
the dataset consisted of 48502 participants. The results showed 85% accuracy, 93% AUC, 90% precision,
78% recall, and 90% specificity. In [19], the prediction model was developed using MLP and SVM
algorithms. In this study, the dataset had 1611 participants. The results showed 75 % accuracy, 74 %
specificity, and 81% AUC. In [21], XGBoost and LightGBM algorithms were used to develop a
prediction model. In this study, the dataset consisted of 34387 participants, and the resulting AUC was
80.4%.</p>
      <p>Accuracy
100
80
60
40
20
0</p>
      <p>Ref.[7]</p>
      <p>Ref.[8]</p>
      <p>Ref.[10] Ref.[11] Ref.[13] Ref.[14] Ref.[15] Ref.[16] Ref.[18] Ref.[19]</p>
      <p>It can be seen from Figure 4 that in almost all of the analyzed articles, results were obtained based
on an accuracy evaluation metric. If we compare the results on accuracy with each other, Ref. [10] had
the highest result and Ref. [8] had the lowest result. The results of Ref. [11] and Ref. [19] are close to
each other. Also, the results of Ref. [10] and Ref. [14] are close to each other.</p>
    </sec>
    <sec id="sec-4">
      <title>4. Conclusion</title>
      <p>The paper analyzes the studies conducted on dataset development to assess the risk of developing
gestational diabetes mellitus. The results of the study suggest that creating a dataset for assessing the
risk of gestational diabetes is a global research topic. The paper also highlights that active research is
being conducted on all continents of the world to create a dataset for gestational diabetes mellitus.</p>
      <p>Based on the results of the analysis, the following will be conclusions:</p>
      <p>Having a large amount of data in a dataset always does not necessarily lead to increased accuracy
in machine-learning models. It is important to note that although having more data can be beneficial,
it is not the only factor that contributes to model accuracy;</p>
      <p>In addition to the dataset, choosing the right prediction models is an important factor in improving
model accuracy;</p>
      <p>Large amounts of irrelevant or noisy data can mislead the model. Data cleaning and feature
engineering are crucial for the effective utilization of large datasets.</p>
      <p>In our future work, we plan to create a dataset of women in Uzbekistan with gestational diabetes,
using the expertise gained from studying leading scientists' remarkable results during the preparation
of this review paper.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <given-names>P.</given-names>
            <surname>Saeedi</surname>
          </string-name>
          et al.,
          <article-title>“Global and regional diabetes prevalence estimates for 2019 and projections for 2030 and 2045: Results from the International Diabetes Federation Diabetes Atlas, 9th edition</article-title>
          ,”
          <source>Diabetes Res. Clin. Pract</source>
          ., vol.
          <volume>157</volume>
          , p.
          <fpage>107843</fpage>
          ,
          <year>2019</year>
          , doi: 10.1016/j.diabres.
          <year>2019</year>
          .
          <volume>107843</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>K.</given-names>
            <surname>Nosirov</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Begmatov</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M.</given-names>
            <surname>Arabboev</surname>
          </string-name>
          , “
          <article-title>Design of a model for multi-parametric health monitoring system</article-title>
          ,
          <source>” International Conference on Information Science and Communications Technologies</source>
          ,
          <string-name>
            <surname>ICISCT</surname>
          </string-name>
          <year>2020</year>
          . pp.
          <fpage>1</fpage>
          -
          <lpage>5</lpage>
          ,
          <year>2020</year>
          . doi:
          <volume>10</volume>
          .1109/ICISCT50599.
          <year>2020</year>
          .
          <volume>9351522</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <given-names>P.</given-names>
            <surname>Bose</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S. K.</given-names>
            <surname>Bandyopadhyay</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Bhaumik</surname>
          </string-name>
          , and
          <string-name>
            <given-names>S.</given-names>
            <surname>Poddar</surname>
          </string-name>
          , “
          <article-title>Female Diabetic Prediction in India Using Different Learning Algorithms</article-title>
          ,” Univers. J. Public Heal., vol.
          <volume>9</volume>
          , no.
          <issue>6</issue>
          , pp.
          <fpage>460</fpage>
          -
          <lpage>471</lpage>
          ,
          <year>2021</year>
          , doi: 10.13189/ujph.
          <year>2021</year>
          .
          <volume>090614</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <given-names>M.</given-names>
            <surname>Arabboev</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Begmatov</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Rikhsivoev</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Saydiakbarov</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Uraimov</surname>
          </string-name>
          , and
          <string-name>
            <given-names>K.</given-names>
            <surname>Nosirov</surname>
          </string-name>
          , “
          <article-title>Gestational diabetes mellitus risk assessment using artificial intelligence: a review,”</article-title>
          <source>in International Conference on Information Science and Communications Technologies: Applications, Trends and Opportunities</source>
          ,
          <year>2023</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <given-names>J.</given-names>
            <surname>Shen</surname>
          </string-name>
          et al.,
          <article-title>“An innovative artificial intelligence-based app for the diagnosis of gestational diabetes mellitus (GDM-AI): Development study</article-title>
          ,”
          <source>J. Med. Internet Res.</source>
          , vol.
          <volume>22</volume>
          , no.
          <issue>9</issue>
          ,
          <string-name>
            <surname>Sep</surname>
          </string-name>
          .
          <year>2020</year>
          , doi: 10.2196/21573.
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <given-names>Z.</given-names>
            <surname>Zhang</surname>
          </string-name>
          et al.,
          <source>“Machine Learning Prediction Models for Gestational Diabetes Mellitus: Metaanalysis,” J. Med. Internet Res.</source>
          , vol.
          <volume>24</volume>
          , no.
          <issue>3</issue>
          ,
          <year>2022</year>
          , doi: 10.2196/26634.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <given-names>R.</given-names>
            <surname>Jader</surname>
          </string-name>
          and
          <string-name>
            <given-names>S.</given-names>
            <surname>Aminifar</surname>
          </string-name>
          , “
          <article-title>Predictive Model for Diagnosis of Gestational Diabetes in the Kurdistan Region by a Combination of Clustering and Classification Algorithms: An Ensemble Approach</article-title>
          ,
          <source>” Appl. Comput. Intell. Soft Comput.</source>
          , vol.
          <year>2022</year>
          ,
          <year>2022</year>
          , doi: 10.1155/
          <year>2022</year>
          /9749579.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <given-names>D.</given-names>
            <surname>Maheshwari</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Garcia-Zapirain</surname>
          </string-name>
          , and
          <string-name>
            <surname>D.</surname>
          </string-name>
          Sierra-Soso, “
          <article-title>Machine learning applied to diabetes dataset using Quantum versus Classical computation</article-title>
          ,”
          <source>2020 IEEE Int. Symp. Signal Process. Inf.</source>
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Technol. ISSPIT</surname>
          </string-name>
          <year>2020</year>
          ,
          <year>2020</year>
          , doi: 10.1109/ISSPIT51521.
          <year>2020</year>
          .
          <volume>9408944</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <given-names>S.</given-names>
            <surname>Jeyaparam</surname>
          </string-name>
          and
          <string-name>
            <given-names>R.</given-names>
            <surname>Agha-Jaffar</surname>
          </string-name>
          , “GDM Dataset,” Figshare,
          <year>2023</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          https://doi.org/10.6084/m9.figshare.
          <volume>21806472</volume>
          .v1
          <string-name>
            <given-names>S.</given-names>
            <surname>Nagarajan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Ramasubramanian</surname>
          </string-name>
          , and R. . Chandrasekaran, “
          <article-title>Data Mining Techniques for Performance Evaluation of Diagnosis in Gestational Diabetes,”</article-title>
          <source>Int. J. Curr. Res. Acad. Rev.</source>
          , vol.
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          2, no.
          <issue>10</issue>
          , pp.
          <fpage>91</fpage>
          -
          <lpage>98</lpage>
          ,
          <year>2014</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <given-names>H.</given-names>
            <surname>Liu</surname>
          </string-name>
          et al., “
          <article-title>Machine learning risk score for prediction of gestational diabetes in early pregnancy in Tianjin</article-title>
          , China,” Diabetes.
          <source>Metab. Res. Rev.</source>
          , vol.
          <volume>37</volume>
          , no.
          <issue>5</issue>
          ,
          <year>2021</year>
          , doi: 10.1002/dmrr.3397.
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <given-names>R. L.</given-names>
            <surname>Lawrence</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C. R.</given-names>
            <surname>Wall</surname>
          </string-name>
          , and
          <string-name>
            <given-names>F. H.</given-names>
            <surname>Bloomfield</surname>
          </string-name>
          , “
          <article-title>Prevalence of gestational diabetes according to commonly used data sources: An observational study,” BMC Pregnancy Childbirth</article-title>
          , vol.
          <volume>19</volume>
          , no.
          <issue>1</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>9</lpage>
          ,
          <year>2019</year>
          , doi: 10.1186/s12884-019-2521-2.
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <given-names>H.</given-names>
            <surname>Qiu</surname>
          </string-name>
          et al.,
          <source>“Electronic Health Record Driven Prediction for Gestational Diabetes Mellitus in Early Pregnancy,” Sci. Rep</source>
          ., vol.
          <volume>7</volume>
          , no.
          <issue>1</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>13</lpage>
          ,
          <year>2017</year>
          , doi: 10.1038/s41598-017-16665-y.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <given-names>N.</given-names>
            <surname>Prema</surname>
          </string-name>
          and
          <string-name>
            <given-names>M.</given-names>
            <surname>Pushpalatha</surname>
          </string-name>
          , “
          <article-title>Analysis of Risk Factors of Gestational Diabetes Mellitus (GDM) Using Data Mining,”</article-title>
          <source>J. Women's Heal. Issues Care</source>
          , vol.
          <volume>8</volume>
          , no.
          <issue>2</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>4</lpage>
          ,
          <year>2018</year>
          , doi: 10.4172/
          <fpage>2325</fpage>
          -
          <lpage>9795</lpage>
          .
          <fpage>1000329</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <given-names>B.</given-names>
            <surname>Pranto</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S. M.</given-names>
            <surname>Mehnaz</surname>
          </string-name>
          ,
          <string-name>
            <given-names>E. B.</given-names>
            <surname>Mahid</surname>
          </string-name>
          ,
          <string-name>
            <given-names>I. M.</given-names>
            <surname>Sadman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Rahman</surname>
          </string-name>
          , and
          <string-name>
            <given-names>S.</given-names>
            <surname>Momen</surname>
          </string-name>
          , “
          <article-title>Evaluating machine learning methods for predicting diabetes among female patients in Bangladesh,” Inf</article-title>
          ., vol.
          <volume>11</volume>
          , no.
          <issue>8</issue>
          ,
          <year>2020</year>
          , doi: 10.3390/INFO11080374.
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <given-names>X.</given-names>
            <surname>Lu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Cai</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Z.</given-names>
            <surname>Xing</surname>
          </string-name>
          , and
          <string-name>
            <given-names>J.</given-names>
            <surname>Huang</surname>
          </string-name>
          , “
          <source>Prediction of Gestational Diabetes and Hypertension Based on Pregnancy Examination Data,” J. Mech. Med</source>
          . Biol., vol.
          <volume>22</volume>
          , no.
          <issue>3</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>18</lpage>
          ,
          <year>2022</year>
          , doi: 10.1142/S0219519422400012.
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>S. D.</surname>
          </string-name>
          Cooray et al., “
          <article-title>Temporal validation and updating of a prediction model for the diagnosis of gestational diabetes mellitus,”</article-title>
          <string-name>
            <given-names>J.</given-names>
            <surname>Clin</surname>
          </string-name>
          . Epidemiol., vol.
          <volume>164</volume>
          , pp.
          <fpage>54</fpage>
          -
          <lpage>64</lpage>
          ,
          <year>2023</year>
          , doi: 10.1016/j.jclinepi.
          <year>2023</year>
          .
          <volume>08</volume>
          .020.
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <given-names>Y.</given-names>
            <surname>Belsti</surname>
          </string-name>
          et al.,
          <article-title>“Comparison of machine learning and conventional logistic regression-based prediction models for gestational diabetes in an ethnically diverse population; the Monash GDM Machine learning model</article-title>
          ,
          <source>” Int. J. Med</source>
          . Inform., vol.
          <volume>179</volume>
          , no.
          <source>September</source>
          , p.
          <fpage>105228</fpage>
          ,
          <year>2023</year>
          , doi: 10.1016/j.ijmedinf.
          <year>2023</year>
          .
          <volume>105228</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <string-name>
            <given-names>G.</given-names>
            <surname>Cubillos</surname>
          </string-name>
          et al.,
          <article-title>“Development of machine learning models to predict gestational diabetes risk in the first half of pregnancy,” BMC Pregnancy Childbirth</article-title>
          , vol.
          <volume>23</volume>
          , no.
          <issue>1</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>19</lpage>
          ,
          <year>2023</year>
          , doi: 10.1186/s12884-023-05766-4.
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <surname>M. Zulueta</surname>
          </string-name>
          et al.,
          <article-title>“Development and validation of a multivariable genotype-informed gestational diabetes prediction algorithm for clinical use in the Mexican population: Insights into susceptibility mechanisms</article-title>
          ,
          <source>” BMJ Open Diabetes Res. Care</source>
          , vol.
          <volume>11</volume>
          , no.
          <issue>2</issue>
          ,
          <year>2023</year>
          , doi: 10.1136/bmjdrc-2022-003046.
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <given-names>B. S.</given-names>
            <surname>Kang</surname>
          </string-name>
          et al.,
          <article-title>“Prediction of gestational diabetes mellitus in Asian women using machine learning algorithms</article-title>
          ,
          <source>” Sci. Rep</source>
          ., vol.
          <volume>13</volume>
          , no.
          <issue>1</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>10</lpage>
          ,
          <year>2023</year>
          , doi: 10.1038/s41598-023-39680-8.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>