<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Architecture of the Platform for Big Data Preprocessing and Processing in Medical Sector</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Eugen Zasoba</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Vladyslav Mykhaikyshyn</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mykhailo Osypov</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Lviv Polytechnic National University</institution>
          ,
          <addr-line>Lviv</addr-line>
          ,
          <country country="UA">Ukraine</country>
        </aff>
      </contrib-group>
      <fpage>0000</fpage>
      <lpage>0001</lpage>
      <abstract>
        <p>The paper presents the architecture of the platform for Big data preprocessing and processing. The method for prevention of the risk disease based on Probabilistic Production Dependencies is developed. The investigation is started from the platform for Big data in medical domain analysis. The platform consists of 6 layers: Data layer, Communication layers, Preprocessing layer, Data processing layer, User layer, Integrational layer. The platform architecture for Big data processing in medical sector is developed. The accuracy of the proposed method is estimated.</p>
      </abstract>
      <kwd-group>
        <kwd>EHealth</kwd>
        <kwd>Probabilistic Production Dependencies</kwd>
        <kwd>Big data</kwd>
        <kwd>Data Mining</kwd>
        <kwd>Medical Sector</kwd>
        <kwd>Big Data Preprocessing</kwd>
        <kwd>Big Data Processing</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>In recent years, the health policies are transforming form acoute intervention driven
towards prevention and self-responsibility of the individuals by implementing
measures for increasing health literacy. Particularly important for the implementation
of the prevention programs are the literacy of young medical staff, the completeness
and timeliness of receiving information about patients, the ability to observe the
patient, not only in the hospital, but also at home. That is why integration of information
technologies, intelligent devices and systems into all spheres of life takes place. It is
important that the technologies used by the patient be created taking into account the
requirements of the universal design, that is, without the need for adaptation or design
changes if used by a person with special needs. The engineering branch is engaged in
the design of "smart" homes, businesses and cities. Such complex intelligent systems
collect process and transmit large amounts of data during their operation, including in
the medical sector. In this process, microcontrollers, communication equipment,
sensors, as well as ambulances, the next branch of a health facility, family doctors, etc.,
process the information coming from patients. Devices may have failures, as well as
errors of the first and second kind, which leads to incomplete or inaccurate
information about the patient, and, consequently, adversely affects the adoption of medical
decisions.</p>
      <p>The aim of the paper is analysis of the existing stage of data mining technics and
Big data approach in EHealth and development of the novel technic for prevention of
the risk disease based on Probabilistic Production Dependencies.
2</p>
    </sec>
    <sec id="sec-2">
      <title>State of art</title>
      <p>
        Data mining is widely used in medicine. Norman et al [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ] is proposed to analyze data
gaps and fill them based on decision trees, EM algorithm, and regression approach to
predict missing data using prediction functions. Similar results were obtained by
Khanmohammadi [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] for associative classification of medical data and Constantinou
[
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] for comprehensive questioning and interviewing of data on Bayesian models to
support medical decision making. The Bayesian network was also used to reformulate
the QMR model in decision theory. However, with the spread of Big data technology,
Bayesian networks were not fast enough. Therefore, Tang et al [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] developed a
method for parallelizing Bayesian networks. Similar results were obtained by Anders
L.Madse [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ]. Bayesian networks are also used to diagnose diseases, such as Lakho
and Seixas [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. However, even in the case of parallelism, it is advisable to use
Bayesian networks in combination with other methods of machine learning for
multiparameter, large-scale and dynamic medical data flows [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ].
      </p>
      <p>
        Currently, Hybrid Systems of Computational Intelligence are widely used to solve
a wide class of Medical Data Mining problems, combining the approximating and
extrapolating properties of artificial neural networks (ANNs), both traditional shallow
ones, and more advanced deep neural networks (DNNs), interpretability of the results
of fuzzy reasoning systems (FRS), the ability to find the best architectures for solving
specific problems, provided by using the apparatus of evolving systems [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ]. Medical
Data Mining considers the most complex tasks of classical Data Mining due to the
fact that the generated clusters are usually nonconvex, overlapped, nonstationary;
initial data are corrupted by “gaps” - omissions and abnormal outliers, the initial
samples to be processed can be short, which leads to undesirable overfitting, as well as
extra-long (Big Data), which leads to the inability to work in batch mode, and
requires the use of Data Stream Mining methods. It is clear that in this situation,
recurrent online data processing methods come to the forefront - training, which allows not
to accumulate huge amounts of information, but to “forget” the data after it has been
processed.
      </p>
      <p>In modified applications, the most problematic are the problems of diagnostics and
early fault detection of controlling biomedical signals, in doing so comes to the fore
not only the accuracy of the results obtained but also the time required to obtain them
come to the fore, which leads to the need to abandon traditional mutiepoch learning.</p>
      <p>Currently, to solve the mentioned problems, the most effective from the accuracy
of the obtained results are deep neural networks, which are however completely
ineffective in the conditions of short learning sets and unsuitable for operating in online
mode, when data are fed to processing in real time.</p>
      <p>
        Probabilistic neural networks (PNN) introduced by D. Specht [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] and trained with
the so-called lazy learning according to principle “neurons in data points” can serve as
an alternative to DNN in diagnostic-recognition tasks. PNN training is very fast
(“just-in-time-models”), however, these networks suffer from “curse of
dimensionality”, and besides, these crisp systems are inefficient in the context of overlapping
classes. In connection with this, it is advisable to develop the evolutionary PNN
architecture, which allows limiting the network dimension to the conditions of the data flow
arriving for processing.
      </p>
      <p>
        Bhatt [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ] developed an approach to collecting medical data using the Internet of
Things and a neural network ensemble to detect unusual human movements.
However, in the case of solving highly specialized problems, which is characteristic of the
field of medicine, the error of training of the neural network ensemble is much higher
than the error of the operation of one network. In addition, the training of the neural
network ensemble is quite time-consuming and costly. In Silva-Ramírez E. L [
        <xref ref-type="bibr" rid="ref11">11</xref>
        ], a
multilayer perceptron method is used to solve the classification and fill in data gaps,
and is based on a combination of multilayer perceptrons and k-nearest neighbors.
      </p>
      <p>Therefore, data mining techniques are used to solve many of the problems of
processing and analyzing medical information. However, there are no comprehensive
studies aimed at identifying the patient's condition without specifying the type of
anamnesis. Specific tasks have been addressed in this area, but studies have only
partially addressed the phenomena of Big Data and the Internet of Things, in-depth
analysis, and visualization of accumulated data to support personalized treatment
decisions.</p>
      <p>
        Big Data has attracted much attention in academia and industry fields in the last
few years. The trend of utilizing medical big data has increased tremendously as well.
The medical big data is generated from medical records, medicine, research, etc. but it
is not well connected due to its long-term development in the last decades in most
hospitals without high-level strategic guide and plan [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ]. They propose a platform
called Data Mid-Platform (DMP) to solve the above problems. The feasibility of this
mid-platform is highly recommended and certified by several domain experts and it is
under construction in a real project.
      </p>
      <p>
        In the context of “Big Data +”, people began to study the application of data
visualization to medical data. Data visualization can make full use of the human sensory
vision system to guide users through data analysis and present information hidden
behind the data in an intuitive and easy-to-use manner. There is the problem of
visualizing such a large amount of data in a way that is understandable to humans: doctors,
patients. In addition, proper visualization allows you to extract general knowledge and
track trends and dependencies. An example of a platform for the visualization of large
medical data sets is presented in [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ].
      </p>
      <p>
        Healthcare (Cardiovascular diseases) has now become a major concern in the
economic as well social regiment of almost everyone. The medical system should be able
to handle long-term and continuous monitoring, inconspicuous monitoring and should
be highly sensitive to the patient's ECG signal. IoT helps the ECG signals to get
transmitted through the sensor via a gateway using communication protocols like
Bluetooth, Zigbee, 3G/4G, Wi-Fi, LAN, etc. Then this data can be sent to numerous
places like the doctor's end, caregiver or the cloud-storage for analysis or processing,
etc. The doctor at the remote location can view the patient's report on various smart
devices, all thanks to the computing technology. This helps in dealing with the
emergencies. Big Data plays a pivotal role in this system as it provides with data analysis,
decision-making, extracting useful information via algorithms, intelligent storage etc
[14. 15]. The combination of these three technologies will make the world a more
secure place to live. Also medical imaging is often used and gives big data medical
records [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. When we are able to analyze all these diagnostic data, it becomes
possible to predict the patient's condition based on previous medical history [
        <xref ref-type="bibr" rid="ref17 ref18">17, 18</xref>
        ]. Table
1 presents comparison of Big data representation.
      </p>
      <p>So, for the healthcare domain all of Big data representation models can be used.</p>
    </sec>
    <sec id="sec-3">
      <title>Methods and results</title>
      <sec id="sec-3-1">
        <title>Big data platform development</title>
        <p>The first layer consists of all possible sources of information. Such approach
allows usage of this platform as part of smart house too. Combination of indoor and
outdoor parameters allows finding dependencies between patient’s conditions and the
weather and using these dependencies for the prediction of the new state of the
patient.</p>
        <p>
          The second layer is used for the communication between the rests of layers. It
is very important to provide the secure transmission and saving of the personal data.
In this paper, we are not considered the security aspects. More details about security
features of medical data is given in [
          <xref ref-type="bibr" rid="ref23">23</xref>
          ].
        </p>
        <p>
          The next layer is preprocessing layer. The medical data is collected from
various sources given at data layer. That is why the filtering, cleaning, wrapping and
missed data imputation methods are used. The next important thing is that medical
data is collected from different countries. Each country has own medical protocols
description written in different languages. In addition, different medical devices are
used in hospitals from different countries. That is why the main tasks of preprocessing
layers are:
 Anonimazation and Data protection – for medical data privacy,
 Missing data recovery [
          <xref ref-type="bibr" rid="ref24 ref25">24, 25</xref>
          ] – for recovery the gaps from devices and medical
records,
 Medical records preprocessing – for semi structural data processing and
comparison of the protocols in different countries,
 Pharmacian data preprocessing [26] – for semi structural data processing and
avoidance of language barriers.
        </p>
        <p>Data processing layer consists of the following modules:
 Online diagnostic based on fuzzy neural networks [27] – for fast training process,
 Personalization of the treatment based on the decision tree [28] – for the most
appropriative medications selection,
 The risk prevention model – see below in more details,
 Statistical model for the risk prediction [29] – for dependencies mining between
paramenets of the pation and the rest of smart home parameters.
 The ontology development [30] – for increasing of the literacy of the young
doctors.</p>
        <p>The user layer represents the different groups of user interested in the platform usage.</p>
        <p>The integration layer represents the list of standards and open platforms related to
the proposed platform.
3.2</p>
      </sec>
      <sec id="sec-3-2">
        <title>Big data platform development</title>
        <p>The usage of a variety of processing technologies of emerging Big Data will allow to
study causal mechanisms of modelling and predicting treatment stages, taking into
account individual peculiarities of the patient, to analyze medicines’ data and find
their key characteristics. This information we use for development of innovative
approaches to improving the methodology of risk stratification, to model the therapeutic
schemes, to improve medical care quality through personalization of treatment
regimens for patients.</p>
        <p>The analysis of patients’ information will help to identify common social
characteristics and thus contribute to the development of recommendations for disease
prevention</p>
        <p>This also will stimulate the healthy people to monitor their health parameters.</p>
        <p>Analyzing large amounts of data requires identifying attribute groups that form
functional dependencies. However, in the real world, data sets are much more
common, with important dependencies defined only on a subset of key attribute group
values; we will call such dependencies partial functional dependencies. The partial
functional dependencies are modified associative rules but for part of the data. The
main idea of this method is to process data given in different structure or
semistructured data. The algorithm for the Probabilistic Production Dependencies building is
built on lazy calculation approach (modification of FP-tree). It allows reducing the
time complexity and using the parallel and distributed mode for calculation.</p>
        <p>Probabilistic Production Dependency (PPD) is the kind of functional dependency
(F-dependency) arized in relational databases. This is also similar dependency to
associative rules built on Aproory algorithm or something similar. The main difference
between associative rules and PPD is that PPD will generate from existing functional
dependencies in dataset. For this purpose, the Armstrong rules are used.</p>
        <p>PPD:    = {  },   ∈  ,  = {  },   ∈  , :  ( ∈  →  ∈  ) =  ,
where k and d are the tuples of groups of attributes K and D respectively.</p>
        <p>The main indicator of the reliability of such a dependency is the ratio of the
number of objects that such PPD has to the number of objects in the selection:
 (PPD) = |  ∈  ∧  ∈ ( )|</p>
        <p>|  ∈ ( )|</p>
        <p>The calculation of the reliability of the implementation of such dependence is
based on the possibility of decomposition of such dependence into components of the
PPD:
(1)
(2)
(3)
 ( ∈  →  ∈  ) = ∑∈   ( ∈  →  =   ) =
∑
  ∈</p>
        <p>∑| =   ∧  =  |</p>
        <p>∑| =  |</p>
        <p>As in the case of F-dependencies (functional dependencies), the set of
classification rules that take place in a given relation can be represented by some subset of
them, from which by means of output rules all classification rules of a given relation
can be obtained. Since classification rules are an extension of F-dependencies, it is
worth considering axiom transformations for functional dependencies for
classification rules.</p>
        <p>Therefore, the algorithm for PPD mining consists of two steps:
1. Frequent pattern mining – based on Apriory algorithm.
2. The existing pattern splitting and decomposition – based on reliability ration.
In the patient data analysis, the sequence of events is often of interest. When detecting
regularities in such sequences, it is possible to predict with some degree the
occurrence of events in the future, which allows us to make more correct decisions. A
sequence is called an ordered set of objects. Using the hierarchy allows you to
determine the connection that goes into higher levels of the hierarchy, since the support for
the set can increase if the entry of the group, and not its object, is counted. In the
hierarchical structure of objects, you can change the nature of the search by changing the
analyzed level. Moving up the hierarchy, we summarize the data and reduce their
number, and vice versa. So, Probabilistic Production Dependencies will be a
combination of associative rules and sequential rules.</p>
        <p>The proposed algorithm makes it possible to assert that the task of detecting
Probabilistic Production Dependencies in distributed databases belongs to the class of
Ptasks. Low asymptotic complexity of the established association rules mining
algorithm and a wide set of data types supported for analysis allow to apply the
established algorithm in practically all subject areas working with association
dependencies in data. Algorithm for finding association dependencies is well-solved with
MapReduce.
4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Results</title>
      <p>The proposed algorithm is tested on open data set for associative riles mining [31].
The data set used for this application is the Adult data set in the Machine Learning
Repository UCI.</p>
      <p>The AOG model is an oriented acyclic graph, where each vertex of a graph
corresponds to a variable with assigned parameters. In Bayesian networks, parameters are
given as local conditional distribution of probabilities of the values of the variables P
(Xi | F (Xi)).</p>
      <p>And in Gaussian networks - as the coefficients of linear equations (for edges) and
dispersion of deviations (for vertices).</p>
      <p>The construction of the AOG-models meets the problem of reproducing the model
from the statistical data. These include methods for restoring the "Collifinder" and
"Proliferator-C" AOG model, generalizing the Chow &amp; Liu method. The use of
Collifinder and Proliferator-C allows recognition of transitive, synergistic, and combined
associations, and thus provides a reliable and effective method for reproducing
structures of single-threaded dependency models without first-level tests.</p>
      <p>The problem of the above-described methods is the need to specify elementary
conditional relationships for constructing a graph of the AOG-model.</p>
      <p>Finding such dependencies goes beyond these algorithms.</p>
      <p>Own algorithm. The paper proposes an algorithm for extracting PPD. For the
relationship with the scheme  = {  ,  (  )},  = ̅1̅̅,̅̅̅, it allows you to find
statistiA
cally significant rules that reflect the dependence of the attribute m on the
utes  1,  2, . . . ,   −1 , that is, the dependence of the species  1,  2, . . . ,   −1 →   .</p>
      <p>As a measure of statistical significance, the Kullbach-Leibler information measure
is used.</p>
      <p>The algorithm allows you to look for only dependencies defined on the whole set
of input data; in addition, it has a high computational complexity if there are many
classification rules.</p>
      <p>The result of own algorithm is given on fig. 2.
If we need to search for associative rules among individual subjects, you can find that
in many cases, associations with high support for individual events are practically
absent. Therefore, in spite of the fact that, in general, event association support1 -&gt;
Event2 can be very high, association support between individual types of events is
likely to be low. Thus, such associations, although they may be of interest, will be
excluded from consideration as they will not meet a certain minimum support
threshold Smin .
longs:</p>
      <p>To solve this problem when seeking associative rules, not individual subjects, but
their hierarchy is considered. If there are no such interesting associations on the lower
hierarchical levels, then they may occur at higher levels. In other words, support for
an individual object will always be less than the support of the group to which it
be ( ) &gt;  (  ),
 ( ) =
∑
 =1 

 ,</p>
      <p>Where I the group is in the hierarchy;   j - this item is included in the given group.
The reasons for this are obvious: the total support for the group is equal to the amount
of support for the items included in it:
where n - the number of items in the group.</p>
      <p>Associative rules found for objects or events located at different hierarchical
levels are called multilevel rules.</p>
      <p>Going down to the lower levels of abstraction, the descendants of only those
categories and subcategories that are frequent sets are analyzed, that is, there are at least a
predetermined number of times, where k - the number of the level.</p>
      <p>Searching for frequent sequences runs from level 1 to the maximum possible. The
results of successive passes will be presented in the Table 2.</p>
      <p>F1</p>
      <p>Thus, the maximum sequences are &lt;1; 2; 3; 4&gt;, &lt;1; 3; 5&gt; and &lt;4; 5&gt; because they
are not contained in sequences of greater length. They will be sought after by
successive templates.
5</p>
    </sec>
    <sec id="sec-5">
      <title>Conclusion</title>
      <p>The study presents the architecture of the platform for Big data preprocessing and
processing. The state of art in this domain is presented. The method for prevention of
the risk disease based on Probabilistic Production Dependencies is developed. The
accuracy of the proposed method is estimated.
26. Bodyanskiy, Y., Tyshchenko, O. (2019). A Hybrid Cascade Neuro–Fuzzy Network with
Pools of Extended Neo–Fuzzy Neurons and its Deep Learning. International Journal of
Applied Mathematics and Computer Science, 29(3), 477-488.
27. Shakhovska, N., Fedushko, S., Shvorob, I., &amp; Syerov, Y. (2019). Development of mobile
system for medical recommendations. Procedia Computer Science, 155, 43-50.
https://doi.org/10.1016/j.procs.2019.08.010
28. Melnykova, N. (2017). Semantic search personalized data as special method of processing
medical information. In Advances in Intelligent Systems and Computing (pp. 315-325).</p>
      <p>Springer, Cham.
29. Fedushko, S., Ustyianovych, T., &amp; Gregus, M. (2020). Real-time high-load infrastructure
transaction status output prediction using operational intelligence and big data
technologies. Electronics, 9(4), 668. https://doi.org/10.3390/electronics9040668
30. Lytvyn, V., Vysotska, V., Dosyn, D., Lozynska, O., Oborska, O. (2018). Methods of
building intelligent decision support systems based on adaptive ontology. In 2018 IEEE Second
International Conference on Data Stream Mining &amp; Processing (DSMP), pp. 145-150.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Norman</surname>
            ,
            <given-names>G. R.</given-names>
          </string-name>
          , et al. (
          <year>2018</year>
          ).
          <article-title>Expertise in medicine</article-title>
          and surgery
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Khanmohammadi</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Adibeig</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Shanehbandy</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          (
          <year>2017</year>
          ).
          <article-title>An improved overlapping kmeans clustering method for medical applications</article-title>
          .
          <source>Expert Systems with Applications</source>
          ,
          <volume>67</volume>
          ,
          <fpage>12</fpage>
          -
          <lpage>18</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Constantinou</surname>
            ,
            <given-names>A. C.</given-names>
          </string-name>
          , et al. (
          <year>2016</year>
          ).
          <article-title>From complex questionnaire and interviewing data to intelligent Bayesian network models for medical decision support</article-title>
          .
          <source>Artificial intelligence in medicine</source>
          ,
          <volume>67</volume>
          ,
          <fpage>75</fpage>
          -
          <lpage>93</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>Y.</given-names>
            <surname>Tang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <surname>K. M. Cooper</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          <string-name>
            <surname>Li</surname>
          </string-name>
          ,
          <article-title>Towards big data Bayesian network learning-an ensemble learning based approach</article-title>
          ,
          <source>International Congress on Big Data</source>
          . pp.
          <fpage>355</fpage>
          -
          <lpage>357</lpage>
          ,
          <year>2014</year>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Kjaerulff</surname>
          </string-name>
          , U. B.,
          <string-name>
            <surname>Madsen</surname>
            ,
            <given-names>A. L.</given-names>
          </string-name>
          (
          <year>2008</year>
          ).
          <article-title>Bayesian networks and influence diagrams</article-title>
          . Springer Science+ Business Media,
          <volume>200</volume>
          ,
          <fpage>114</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Lakho</surname>
          </string-name>
          ,
          <string-name>
            <surname>Shamshad</surname>
          </string-name>
          , et al.
          <article-title>Decision Support System for Hepatitis Disease Diagnosis using Bayesian Network</article-title>
          .
          <source>Sukkur IBA Journal of Computing and Mathematical Sciences</source>
          ,
          <year>2017</year>
          ,
          <volume>1</volume>
          .2:
          <fpage>11</fpage>
          -
          <lpage>19</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Seixas</surname>
            ,
            <given-names>Flávio</given-names>
          </string-name>
          <string-name>
            <surname>Luiz</surname>
          </string-name>
          , et al.
          <article-title>A Bayesian network decision model for supporting the diagnosis</article-title>
          ,
          <year>2014</year>
          ,
          <volume>51</volume>
          :
          <fpage>140</fpage>
          -
          <lpage>158</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Bodyanskiy</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dolotov</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>Hybrid systems of computational intelligence evolved from self-learning spiking neural network</article-title>
          .
          <source>Methods and Instruments of Artificial Intelligence</source>
          ,
          <volume>17</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Specht</surname>
            ,
            <given-names>D. F.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Romsdahl</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          (
          <year>1994</year>
          ).
          <article-title>Experience with adaptive probabilistic neural networks and adaptive general regression neural networks</article-title>
          .
          <source>International Conference on Neural Networks</source>
          , Vol.
          <volume>2</volume>
          , pp.
          <fpage>1203</fpage>
          -
          <lpage>1208</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Bhatt</surname>
          </string-name>
          , Chintan; Dey, Nilanjan; Ashour, Amira S. (ed.).
          <article-title>Internet of things and big data technologies for next generation healthcare</article-title>
          .
          <source>2017</source>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11. E. L.
          <string-name>
            <surname>Silva-Ramírez</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          <string-name>
            <surname>Pino-Mejías</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <article-title>López-Coello, “Single imputation with multilayer perceptron and multiple imputation combining multilayer perceptron and k-nearest neighbours for monotone patterns”</article-title>
          ,
          <source>Applied Soft Computing</source>
          , vol.
          <volume>29</volume>
          , pp.
          <fpage>65</fpage>
          -
          <lpage>74</lpage>
          ,
          <year>2015</year>
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <given-names>D.</given-names>
            <surname>Li</surname>
          </string-name>
          , et al.
          <article-title>Practical Data Mid-Platform Design and Implementation for Medical Big Data</article-title>
          ,
          <source>Advanced Information Technology, Electronic and Automation Control Conference (IAEAC)</source>
          ,
          <year>China</year>
          ,
          <year>2019</year>
          , pp.
          <fpage>1042</fpage>
          -
          <lpage>1045</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Yang</surname>
            and
            <given-names>T.</given-names>
          </string-name>
          <string-name>
            <surname>Chen</surname>
          </string-name>
          ,
          <source>Analysis and Visualization Implementation of Medical Big Data Resource Sharing Mechanism Based on Deep Learning</source>
          , vol.
          <volume>7</volume>
          , pp.
          <fpage>156077</fpage>
          -
          <lpage>156088</lpage>
          ,
          <year>2019</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <given-names>M.</given-names>
            <surname>Bansal</surname>
          </string-name>
          and
          <string-name>
            <given-names>B.</given-names>
            <surname>Gandhi</surname>
          </string-name>
          ,
          <article-title>"IoT &amp; Big Data in Smart Healthcare (ECG Monitoring)</article-title>
          ,
          <source>" 2019 International Conference on Machine Learning, Big Data, Cloud and Parallel Computing (COMITCon)</source>
          , Faridabad, India,
          <year>2019</year>
          , pp.
          <fpage>390</fpage>
          -
          <lpage>396</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Suinesiaputra</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          <string-name>
            <surname>Medrano-Gracia</surname>
            ,
            <given-names>B. R.</given-names>
          </string-name>
          <string-name>
            <surname>Cowan</surname>
            and
            <given-names>A. A.</given-names>
          </string-name>
          <string-name>
            <surname>Young</surname>
          </string-name>
          ,
          <article-title>"Big Heart Data: Advancing Health Informatics Through Data Sharing in Cardiovascular Imaging,"</article-title>
          <source>in IEEE Journal of Biomedical and Health Informatics</source>
          , vol.
          <volume>19</volume>
          , no.
          <issue>4</issue>
          , pp.
          <fpage>1283</fpage>
          -
          <lpage>1290</lpage>
          ,
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <given-names>S.</given-names>
            <surname>Xu</surname>
          </string-name>
          , et al.
          <article-title>Cardiovascular risk prediction method based on test analysis and data mining ensemble system</article-title>
          ,
          <source>International Conference on Big Data Analysis</source>
          ,
          <year>2016</year>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>5</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Ruano M. G.</surname>
          </string-name>
          , et al,
          <article-title>Reliability of medical databases for the use of real word data and data mining techniques for cardiovascular diseases progression in diabetic patients</article-title>
          , Global Medical Engineering Physics Exchanges/Pan American Health Care Exchanges, Porto,
          <year>2018</year>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>6</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Melnykova</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Shakhovska</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Melnykov</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          (
          <year>2019</year>
          ).
          <article-title>Using Big Data for Formalization the Patient's Personalized Data</article-title>
          .
          <source>Procedia Computer Science</source>
          ,
          <volume>155</volume>
          ,
          <fpage>624</fpage>
          -
          <lpage>629</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Maté</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Llorens</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>de Gregorio</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2012</year>
          ,
          <article-title>October)</article-title>
          .
          <article-title>An integrated multidimensional modeling approach to access big data in business intelligence platforms</article-title>
          .
          <source>In International Conference on Conceptual Modeling</source>
          (pp.
          <fpage>111</fpage>
          -
          <lpage>120</lpage>
          ). Springer, Berlin, Heidelberg.
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Veerapaneni</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Co-Reyes</surname>
            ,
            <given-names>J. D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chang</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Janner</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Finn</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wu</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , ... &amp;
          <string-name>
            <surname>Levine</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          (
          <year>2020</year>
          , May).
          <article-title>Entity abstraction in visual model-based reinforcement learning</article-title>
          .
          <source>In Conference on Robot Learning</source>
          (pp.
          <fpage>1439</fpage>
          -
          <lpage>1456</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Jagadish</surname>
            ,
            <given-names>H. V.</given-names>
          </string-name>
          , et al (
          <year>2014</year>
          ).
          <article-title>Big data and its technical challenges</article-title>
          .
          <source>Communications of the ACM</source>
          ,
          <volume>57</volume>
          (
          <issue>7</issue>
          ),
          <fpage>86</fpage>
          -
          <lpage>94</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Yang</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhou</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yang</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Zhang</surname>
            ,
            <given-names>Q.</given-names>
          </string-name>
          (
          <year>2020</year>
          ).
          <article-title>A new prediction method for recommendation system based on sampling reconstruction of signal on graph</article-title>
          .
          <source>Expert Systems with Applications</source>
          ,
          <volume>113587</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Pirbhulal</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , et al (
          <year>2019</year>
          ).
          <article-title>A joint resource-aware and medical data security framework for wearable healthcare systems</article-title>
          .
          <source>Future Generation Computer Systems</source>
          ,
          <volume>95</volume>
          ,
          <fpage>382</fpage>
          -
          <lpage>391</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Tkachenko</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Izonin</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kryvinska</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dronyuk</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zub</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          (
          <year>2020</year>
          ).
          <article-title>An Approach towards Increasing Prediction Accuracy for the Recovery of Missing IoT Data based on the GRNN-SGTM Ensemble</article-title>
          . Sensors,
          <volume>20</volume>
          (
          <issue>9</issue>
          ),
          <fpage>2625</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Shakhovska</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vovk</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kryvenchuk</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>Uncertainty reduction in Big data catalogue for information product quality evaluation</article-title>
          ,
          <source>Eastern-European Journal of Enterprise Technologies</source>
          ,Vol
          <volume>1</volume>
          ,
          <source>No</source>
          <volume>2</volume>
          (
          <issue>91</issue>
          ),
          <fpage>12</fpage>
          -
          <lpage>20</lpage>
          (
          <year>2018</year>
          ).
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>