<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Neural Network Classification of Large-sized Multi-segmented Polygon with Formation of Features by the Hilbert-Huang Transform of Hyperspectral Data</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Evgeny S. Nezhevenko</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Institute of Automation and Electrometry SB RAS</institution>
          ,
          <addr-line>Novosibirsk</addr-line>
        </aff>
      </contrib-group>
      <abstract>
        <p>A method of classification of hyperspectral images of terrain is described. It includes preprocessing of the spectral data in the form of conversion to the principal components, spatial processing consisting of finding the empirical mode of the principal components (Hilbert-Huang transform), and classification itself using RBF neural networks. Experiments are conducted on a large-format (580x580) hyperspectral image, and difficult-to-distinguish classes are not united. .</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>Classification method As is known from the pattern recognition theory, object
preprocessing is an extremely important stage of the classification procedure. Here we consider a
spectral-spatial object; hence, preprocessing is performed in the spectral-spatial domain. As it was
already mentioned, each pixel is characterized by 200 spectral features. The analysis shows that
they are mutually correlated to a large extent; therefore, it is logical to pass to non-correlated
features. This can be done by several methods, but the most effective option is to perform
transformation to the principal components (PCs). An important issue is the number of components
to be left in order to lose the minimum possible amount of information. Kaiser and scree tests are
applied for this purpose. After the transformation to the principal components is performed, a
spatial transformation is applied. It was shown in many publications [5, 6, 7] that the use of these
tests ensures significant enhancement of the overall accuracy, which is the percentage of the ratio of
the number of correctly classified pixels to the total number of pixels in the sample. In these
publications, similar to many others, however, the spatial transformation is reduced to changing the
pixel value depending on its neighborhood. The structure of the classified fragment regions is
absolutely ignored, and large errors of classification are observed in the case of
difficult-todistinguish regions. Because of that, in particular, in [6, 7, 8] all soya regions with different
methods of soil treatment were united into one class. The same refers to corn. Our method is based
on another procedure. Each principal component is expanded into the so-called empirical functions
(or intrinsic mode functions, IMFs). In contrast to the Fourier transform or wavelet transform, IMFs
are not defined analytically and are determined exclusively by the analyzed sequence itself. The
basis functions of the transform are formed adaptively, directly from the input data.</p>
      <p>The algorithm of expansion into IMFs is based on constructing smooth envelopes on the
basis of the extreme (maximum and minimum) points of the sequence and subsequent subtraction of
the mean value of the envelopes from the initial sequence. For this purpose, the maximum and
minimum points are found and approximated by splines. These splines are the upper and lower
envelopes. The process of envelope construction is illustrated in Fig. 2.</p>
      <p>The analyzed sequence is presented in Fig. 2 by the thin blue curve. The maximum and
minimum points of this sequence are marked by the red and blue colors, respectively. The
envelopes are shown by the green curve. The mean value is calculated on the basis of two envelopes
(shown by the dashed curve in Fig. 2). The thus-found mean value is further subtracted from the
initial sequence.</p>
      <p>
        These steps are performed to find the first approximation of the sought IMF. For definite
identification of the IMF, it is necessary to find the maximum and minimum points of this IMF
estimate and repeat these steps again. This process (called sieving) is continued until a threshold
condition is satisfied. If the sieving process is successfully finalized, we obtain the first IMF. To
find the next IMF, it is necessary to subtract the found IMF from the initial signal and repeat the
procedure again. The procedure is repeated until all IMFs are found. The algorithm of expansion of
the two-dimensional signal into IMFs formally does not differ from expansion of the
onedimensional signal, thought there are certainly some specific features associated with the search for
extreme points and interpolation [
        <xref ref-type="bibr" rid="ref5">9,10</xref>
        ]. After m intrinsic mode functions of n principal components
are found, each HIS pixel is described by an mxn-dimensional vector of features. All HSI vectors
together with information about the classes are fed to the neural network. Pixels of each class are
randomly divided into three samples: learning sample (LS), control sample (CS), and test sample
(TS). The process of learning is terminated on the basis of results of the CS. The classification
accuracy is checked for all samples, but special care is applied for the TS because it is the TS that
determines the neural network capability to generalization.
      </p>
      <p>Experimental results The HSI shown in Fig. 1a was expanded into the principal
components from which four components were chosen on the basis of the scree test (which include
99.42% of data dispersion). Then each of the PCs was expanded into five IMFs. The first PC and its
five IMFs are shown in Fig.3.</p>
    </sec>
    <sec id="sec-2">
      <title>Fig.3 First PC and its five IMFs</title>
      <p>Thus, after all these transformations, the HSI is described by a 580х580х20 array divided
into 33 classes. The number of pixels in each class is divided in the ratio LS:CS:TS=
50%:25%:25%. Learning is performed in the RBF neural network. The learning time is several
hours. Let us consider the results of classification after the learning process. Let us first give the list
of all classes with numeration that will be used in the tables.</p>
      <p>The classification results are summarized in Table 2. The network architecture in the left
box of the table means 20 input features (number of neurons in the first layer), 851 neurons (RBF
functions) in the hidden layer, and 33 neurons in the output layer (number of classes). The
classification accuracy (CA) is not that high as in experiments performed with smaller fragments
and a smaller number of classes [2, 3], but it should be again recalled that the present study involves
difficult-to-distinguish classes of corn and soya, which are not united into two classes.</p>
    </sec>
    <sec id="sec-3">
      <title>Network architecture</title>
      <p>RBF 20-851-33</p>
    </sec>
    <sec id="sec-4">
      <title>File “All regions of Classification accuracy Learning sample 89.35</title>
      <p>class 33”
Classification accuracy
Control sample
87.36</p>
    </sec>
    <sec id="sec-5">
      <title>Classification accuracy</title>
      <p>Test sample
87.31
Let us consider the results of TS classification in more detail with indication of CA percentage
(Table 3).</p>
      <p>It is seen that the CA percentage varies from 42% (class 3 “Buildings”) to 99.3% (class 58
“Forest”). It should be noted that higher CA values are normally provided by classes with a large
number of pixels: the mean percentage of correct classification is 83.63% based on classes and
87.31% based on pixels.</p>
      <p>Let us consider Table 4, which is the error matrix showing the classes to which erroneously
classified data are referred to. Each row of the matrix is normalized to the number of correctly
classified pixels in the corresponding classes.</p>
      <p>The gray regions are the classes of one plant type: all classes of corn (left) and all classes of
soya (right). It can be concluded from the comparison of these regions with others that the total
cross error in the soya region is several times smaller than that in other regions. It fact testifies that
soya regions with different types of soil processing and other characteristics are indeed hard to
distinguish. This phenomenon for the corn regions is observed to a smaller extent. Probably, this
difference is caused by specific aspects of processing of corn and soya regions.</p>
      <p>Conclusions Thus, the experiments aimed at classification of a large-format hyperspectral
image confirmed the high efficiency of the method including preprocessing in the form of the
transformation to the principal components of the spectral components, spatial transformation in the
form of expansion of the principal components into intrinsic mode functions, and classification in
the neural network. The method ensures a high probability of correct classification of hyperspectral
images having similar spectral compositions and effectively operates in situations with terrain areas
that can be hardly distinguished by conventional methods.</p>
      <p>This work was supported by the Presidium of the Russian Academy of Sciences within the
framework of the Complex Program of Basic Research of the Siberian Branch of the Russian
Academy of Sciences (Project No. II.1.37).</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Review</surname>
            ,
            <given-names>Optoelectron.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Instrum</surname>
            <given-names>.</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>Data</given-names>
            <surname>Process</surname>
          </string-name>
          .
          <year>2018</year>
          , Vol.
          <volume>54</volume>
          , No..
          <volume>6</volume>
          , pp.
          <fpage>582</fpage>
          -
          <lpage>599</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>E.S.</given-names>
            <surname>Nezhevenko</surname>
          </string-name>
          and
          <string-name>
            <given-names>A.S.</given-names>
            <surname>Feoktistov</surname>
          </string-name>
          ,
          <article-title>Investigation of the efficiency of neural network classification with the use of the Hilbert-Huang transform</article-title>
          ,
          <source>in: Proc. Intern</source>
          . Congress “Interexpo
          <string-name>
            <surname>Geo-Sibir</surname>
            <given-names>”</given-names>
          </string-name>
          ,
          <source>April 18-22</source>
          ,
          <year>2016</year>
          , Vol.
          <volume>1</volume>
          , pp.
          <fpage>60</fpage>
          -
          <lpage>64</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <given-names>E.S.</given-names>
            <surname>Nezhevenko</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.S.</given-names>
            <surname>Feoktistov</surname>
          </string-name>
          , and
          <string-name>
            <given-names>O.</given-names>
            <surname>Yu</surname>
          </string-name>
          . Dashevskii,
          <article-title>Neural network classification of hyperspectral images on the basis of the Hilbert-Huang transform</article-title>
          , Optoelectron.,
          <string-name>
            <surname>Instrum</surname>
            <given-names>.</given-names>
          </string-name>
          ,
          <source>Data Proc. 2017</source>
          , Vol.
          <volume>53</volume>
          , No.
          <issue>2</issue>
          , pp
          <fpage>165</fpage>
          -
          <lpage>170</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>M.F. Baumgardner</surname>
            ,
            <given-names>L.L.</given-names>
          </string-name>
          <string-name>
            <surname>Biehl</surname>
            , and
            <given-names>D.A.</given-names>
          </string-name>
          <string-name>
            <surname>Landgrebe</surname>
          </string-name>
          , 220
          <source>Band AVIRIS Hyperspectral Image Data Set: June</source>
          <volume>12</volume>
          , 1992
          <string-name>
            <given-names>Indian</given-names>
            <surname>Pine</surname>
          </string-name>
          <article-title>Test Site 3</article-title>
          . Purdue University Research Repository.
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <source>doi:10</source>
          .4231/R7RX991C.
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <given-names>A.V.</given-names>
            <surname>Kuznetsov</surname>
          </string-name>
          and
          <string-name>
            <given-names>V.V.</given-names>
            <surname>Myasnikov</surname>
          </string-name>
          ,
          <article-title>Comparison of algorithm of controlled elemental classification of hyperspectral images</article-title>
          ,
          <source>Komp. Optika</source>
          ,
          <year>2014</year>
          , Vol.
          <volume>38</volume>
          , No.
          <issue>4</issue>
          , pp.
          <fpage>494</fpage>
          -
          <lpage>502</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>S.M. Borzov</surname>
            ,
            <given-names>A.O.</given-names>
          </string-name>
          <string-name>
            <surname>Potaturkin</surname>
            ,
            <given-names>O.I.</given-names>
          </string-name>
          <string-name>
            <surname>Potaturkin</surname>
          </string-name>
          ,
          <article-title>and</article-title>
          <string-name>
            <given-names>A.M.</given-names>
            <surname>Fedotov</surname>
          </string-name>
          ,
          <article-title>Analysis of the efficiency of classification of hyperspectral satellite images of natural and man-made areas</article-title>
          , Optoelectron.,
          <string-name>
            <surname>Instrum</surname>
            <given-names>.</given-names>
          </string-name>
          ,
          <source>Data Proc., 2016</source>
          , Vol.
          <volume>52</volume>
          , No.
          <issue>1</issue>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>10</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <given-names>S.M.</given-names>
            <surname>Borzov</surname>
          </string-name>
          and
          <string-name>
            <surname>O.I. Potaturkin</surname>
          </string-name>
          ,
          <article-title>Efficiency of the spectral-spatial classification of hyperspectral imaging data, Optoelectron</article-title>
          .,
          <string-name>
            <surname>Instrum</surname>
            <given-names>.</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>Data</given-names>
            <surname>Process</surname>
          </string-name>
          .,
          <year>2017</year>
          , Vol.
          <volume>53</volume>
          , No.
          <issue>1</issue>
          , pp.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <given-names>S.M.</given-names>
            <surname>Borzov</surname>
          </string-name>
          and
          <string-name>
            <surname>O.I. Potaturkin</surname>
          </string-name>
          ,
          <article-title>Classification of hyperspectral images with different methods of training set formation, Optoelectron</article-title>
          .,
          <string-name>
            <surname>Instrum</surname>
            <given-names>.</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>Data</given-names>
            <surname>Process</surname>
          </string-name>
          .,
          <year>2018</year>
          , Vol.
          <volume>54</volume>
          , No.
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          1, pp.
          <fpage>76</fpage>
          -
          <lpage>82</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <given-names>L.</given-names>
            <surname>Vincent</surname>
          </string-name>
          ,
          <article-title>Morphological grayscale reconstruction in image analysis: Applications and efficient algorithms</article-title>
          ,
          <source>IEEE Trans. Image Process., 1993</source>
          , Vol.
          <volume>2</volume>
          , No.
          <issue>2</issue>
          , pp.
          <fpage>176</fpage>
          -
          <lpage>201</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>J.C.</given-names>
            <surname>Carr</surname>
          </string-name>
          ,
          <string-name>
            <given-names>W.R.</given-names>
            <surname>Fright</surname>
          </string-name>
          , and
          <string-name>
            <given-names>R.K.</given-names>
            <surname>Beatson</surname>
          </string-name>
          ,
          <article-title>Surface interpolation with radial basis functions for medical imaging</article-title>
          ,
          <source>Proc. of the Conf. ACM SIGGRAPH</source>
          . Los Angeles, USA,
          <year>August</year>
          12-
          <issue>17</issue>
          ,
          <year>2001</year>
          , pp.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>