<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Comparison of Recent Machine Learning Techniques for Gender Recognition from Facial Images</article-title>
      </title-group>
      <contrib-group>
        <aff id="aff0">
          <label>0</label>
          <institution>Computer Science Department Computer Science Department Computer Science Department Computer Science Department Central Washington University Central Washington University Central Washington University Central Washington University Ellensburg</institution>
          ,
          <addr-line>WA, USA Ellensburg, WA, USA Ellensburg, WA, USA Ellensburg, WA</addr-line>
          ,
          <country>USA and</country>
          <institution>Electronics and Computers Department Transilvania University Bras ̧ov</institution>
          ,
          <country country="RO">Romania</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2016</year>
      </pub-date>
      <fpage>97</fpage>
      <lpage>102</lpage>
      <abstract>
        <p>Recently, several machine learning methods for gender classification from frontal facial images have been proposed. Their variety suggests that there is not a unique or generic solution to this problem. In addition to the diversity of methods, there is also a diversity of benchmarks used to assess them. This gave us the motivation for our work: to select and compare in a concise but reliable way the main state-of-the-art methods used in automatic gender recognition. As expected, there is no overall winner. The winner, based on the accuracy of the classification, depends on the type of benchmarks used.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>A major goal of computer vision and artificial intelligence is
to build computers that can understand or classify concepts
such as gender in the same way humans do.</p>
      <p>Automatic classification of gender from frontal face
images taken under contrived conditions has been well studied
with impressive results. The variety of methods published
in the literature show that there is not a unique or generic
solution to the gender classification problem.</p>
      <p>Applications of gender classification include, image
search, automatic annotation of images, security systems,
face recognition, and real time image acquisition on smart
phones and mobile devices.</p>
      <p>
        The state-of-the-art gender classification methods
generally fall into the following main categories:
Convolutional Neural Networks (CNN), Dual Tree Complex Wavelet
Transform (DTCWT) + a Support Vector Machine (SVM)
classifier, and feature extraction techniques such as Principal
Component Analysis (PCA), Histograms of Oriented
Gradients (HOG) and others with a classifier (SVM, kNN, etc).
The SVM approach is natural, since we have a two class
problem. The CNN is related to the well-known deep
learning paradigm. The DTCWT provides approximate shift
invariance and directionally selective filters (properties
lacking in the traditional wavelet transform) while preserving
the usual properties of perfect reconstruction and
computational efficiency with good well-balanced frequency
responses
        <xref ref-type="bibr" rid="ref14">(Kingsbury 2001)</xref>
        .
      </p>
      <p>To assess gender classification techniques, two types of
benchmarks may be used: standard posed datasets (with well
defined backgrounds, lighting and photographic
characteristics) and datasets containing “In the wild” images that
exhibit the diversity of subjects, settings, and qualities typical
of everyday scenes.</p>
      <p>
        The diversity of the methods and benchmarks makes
a comparison between gender classification a challenging
task, and this gave us the motivation for our work. We
compare state-of-the-art methods used in automatic gender
recognition on two benchmarks: the most popular standard
dataset Facial Recognition Technology (FERET)
        <xref ref-type="bibr" rid="ref20">(Phillips et
al. 2000)</xref>
        and a more challenging data set of “in the wild”
images (Adience)
        <xref ref-type="bibr" rid="ref13 ref9">(Eidinger, Enbar, and Hassner 2014)</xref>
        .
      </p>
      <p>
        We only compare the accuracy of the classification and
not other performance measures (precision, recall, F1 score,
etc). The main reason is that the misclassification cost in this
particular problem is the same, regardless if we misclassify a
male or a female. We also do not compare the running time,
since the experiments are performed on different computer
architectures (the CNN is implemented on a GPU).
Classifiers such as SVMs and feedforward NNs are often
used to classify images after the faces have been cropped out
from the rest of the image, and possibly aligned and
normalized. Various feature extraction methods such as Principal
Component Analysis (PCA), independent component
analysis, Fischer linear discriminants
        <xref ref-type="bibr" rid="ref2">(Belhumeur, Hespanha, and
Kriegman 1997)</xref>
        (Wu et al. 2015), and edge detection
algorithms can be used to encode useful information from
the image that is fed into the classifier, leading to high
levels of accuracy on many benchmarks. Other approaches use
hand-crafted template features to find facial keypoints such
as nose, eyes etc, while also using edge detection methods
(Sobel) and line intensities to separate facial edges from
wrinkles. The resulting feature information, when fed into
a feedforward neural network, allows age and gender to be
classified with overall 85% accuracy on two test sets with
a total of 172 images in the FERET and FGNET databases
        <xref ref-type="bibr" rid="ref13">(Kalansuriya and Dharmaratne 2014)</xref>
        .
      </p>
      <p>
        LDA (Linear Discriminant Analysis) based approaches to
the face recognition task promise invariance to differing
illuminations
        <xref ref-type="bibr" rid="ref2">(Belhumeur, Hespanha, and Kriegman 1997)</xref>
        .
This has been further studied in (Bekios-Calfa,
Buenaposada, and Baumela 2011). Fisher linear discriminant
maximizes the ratio of between- class scatter to that of
withinclass scatter. Independent component analysis has been used
on a small subset (500 images) of the FERET dataset,
leading to 96% accuracy with an SVM classifier
        <xref ref-type="bibr" rid="ref12 ref6">(Jain, Huang,
and Fang 2005)</xref>
        . Likewise, PCA has been used in
conjunction with a genetic algorithm that eliminated potentially
unnecessary features. The remaining features were then fed
to a feedforward neural network for training, and an
overall 85% accuracy was obtained over 3 data sets
        <xref ref-type="bibr" rid="ref23">(Sun et al.
2002)</xref>
        . Various information theory based metrics were also
fused together to produce 99.13% gender classification
accuracy on the FERET
        <xref ref-type="bibr" rid="ref19">(Perez et al. 2012)</xref>
        . To overcome the
challenge of inadequate contrast among facial features using
histogram analysis, Haar wavelet transformation and
Adaboost learning techniques have been employed, resulting in
a 97.3% accuracy on the Extended Yale face database which
contains 17 subjects under 576 viewing conditions
        <xref ref-type="bibr" rid="ref13 ref15">(Laytner,
Ling, and Xiao 2014)</xref>
        . Another experiment describes how
various transformations, such as noise and geometric
transformations, were fed in combination into a series of RBFs
(Radial Basis Functions). RBF outputs were forwarded into
a symbolic decision tree that outputs gender and ethnic class.
94% classification accuracy was obtained using the hybrid
architecture on the FERET database
        <xref ref-type="bibr" rid="ref10">(Gutta, Wechsler, and
Phillips 1998)</xref>
        .
      </p>
      <p>
        HOG (Histogram of Oriented Gradients) is commonly
used as a global feature extraction technique that expresses
information about the directions of curvatures of an image.
HOG features can capture information about local edge and
gradient structures while maintaining degrees of invariance
to moderate changes in illumination, shadowing, object
location, and 2D rotation. HOG descriptors, combined with
SVM classifiers, can be used as a global feature extraction
mechanism (Torrione et al. 2014), while HOG descriptors
can be used on locations indicated by landmark-finding
software in areas such as facial expression classification
        <xref ref-type="bibr" rid="ref7">(De´niz
et al. 2011)</xref>
        . One useful application of variations in HOG
descriptors is the automatic detection of pedestrians, which is
made easier in part because of their predominantly upright
pose
        <xref ref-type="bibr" rid="ref6">(Dalal and Triggs 2005)</xref>
        . In addition, near perfect
results were obtained in facial expression classification when
HOG descriptors were used to extract features from faces
that were isolated through face-finding software
        <xref ref-type="bibr" rid="ref4">(Carcagn`ı
et al. 2015)</xref>
        .
      </p>
      <p>A recent technique proposed for face recognition is the
DTCWT, due to its ability to improve operation under
varying illumination and shift conditions when compared to
Gabor Wavelets and DWT (Discrete Wavelet Transform). The
Extended Yale B and AR face databases were used,
containing a total 16128 images of 38 human subjects under 9 poses
and 64 illumination conditions. It achieved 98%
classification accuracy in the best illumination condition, while low
frequency subband image at scale one (L1) achieved 100%
(Sultana et al. 2014).</p>
      <p>
        Recent years have seen great success in image related
problems through the use of CNN, thereby seeing the
proliferation of a scalable and more or less universal algorithmic
approach to solving general image processing problems, if
enough training data is available. CNNs have had a great
deal of success in dealing with images of subjects and
objects in natural non-contrived settings, along with handling
the rich diversity that these images entail. One investigation
of CNN fundamentals involved training a CNN to classify
gender on images collected on the Internet. 88%
classification accuracy was achieved after incorporating L2
regularization into training, and filters were shown to respond to
the same features that neuroscientists have identified as
fundamental cues humans use in gender classification
        <xref ref-type="bibr" rid="ref13">(Verma
and Vig 2014)</xref>
        . Another experiment
        <xref ref-type="bibr" rid="ref16">(Levi and Hassner 2015)</xref>
        uses a convolutional neural network on the Adience dataset
for gender and age recognition. They used data
augmentation and face cropping to achieve 86% accuracy for gender
classification. This is the only paper we know of that uses
CNN on Adience.
      </p>
      <p>
        A method recently proposed by
        <xref ref-type="bibr" rid="ref13 ref9">(Eidinger, Enbar, and
Hassner 2014)</xref>
        uses an SVM with dropout, a technique
inspired from newer deep learning methods, that has shown
promise for age and gender estimation. Dropout involves
dropping a certain percent of features randomly during
training. They also introduce the Adiance dataset to fulfill the
need for a set of realistic labeled images for gender and
age recognition in quantities needed to prevent overfitting
and allow true generalization
        <xref ref-type="bibr" rid="ref13 ref9">(Eidinger, Enbar, and Hassner
2014)</xref>
        .
      </p>
      <p>As we can see, most of the state-of-the-art methods for
gender classification fall into the categories described in
Section .</p>
    </sec>
    <sec id="sec-2">
      <title>Data sets</title>
      <p>
        A number of databases exist that can be used to benchmark
gender classification algorithms. Most image sets that
contain gender labels suffer from insufficient size, and because
of this we chose two of the larger publicly available datasets:
Color-FERET
        <xref ref-type="bibr" rid="ref20">(Phillips et al. 2000)</xref>
        and Adience
        <xref ref-type="bibr" rid="ref13 ref9">(Eidinger,
Enbar, and Hassner 2014)</xref>
        .
      </p>
      <p>Color FERET Version 2 was collected between
December 1993 and August 1996 and made freely available with
the intent of promoting the development of face recognition
algorithms. Images in the FERET Color database are 512
by 768 pixels and are in PPM format. They are labeled with
gender, pose, name, and other useful labels.</p>
      <p>Although FERET contains a large number of high
quality images in different poses and with varying face
obstructions (beards, glasses, etc), they all have certain similarities
in quality, background, pose, and lighting which make them
very easy for modern machine learning methods to correctly
classify. We used all 11338 images in FERET for which
gender labels exist in our experiments.</p>
      <p>As machine learning algorithms are increasingly used to
process images of varying quality with vast differences in
scale, obstructions, focus, and which are often acquired with
consumer devices such as web cams or cellphones,
benchmarks such as FERET have become less useful. To address
this issue, datasets such as LWS (labeled faces in the wild)
and most recently Adience have emerged. LWS lacks
gender labels but it has accurate names from which gender can
often be deduced automatically with reasonable accuracy.</p>
      <p>
        Adience is a recently released benchmark that contains
gender and approximate age labels separated into 5 folds to
allow duplication of results published by the database
authors. It was created by collecting Flickr images and is
intended to capture all variations of pose, noise, lighting, and
image quality. Each image is labeled with age and gender.
It is designed to mimic the challenges of ”real world”
image classification tasks where faces can be partly obscured
or even partly overlapping (for example in a crowd or when
an adult is holding a child and both are looking at the
camera). Eidlinger, et al published a paper where they used a
SVM and filtering to classify age and gender along with the
release of Adience
        <xref ref-type="bibr" rid="ref13 ref9">(Eidinger, Enbar, and Hassner 2014)</xref>
        .
      </p>
      <p>We used all 19370 of the aligned images from Adience
that had gender labels, to create our training and testing sets
for all experiments that used Adience.</p>
      <p>Using the included labels and meta data in the FERET and
Adience datasets, we generated two files containing reduced
size 51x51 pixel data with values normalized between 0 and
1, followed by a gender label. We choose to resize the
images to 51x51 because this produced the best quality images
after Anti-Aliasing</p>
    </sec>
    <sec id="sec-3">
      <title>Classification methods</title>
      <p>We compare the accuracy of CNN and several SVM based
classifiers. We limit ourselves to methods involving these
two approaches because they are among the most effective
and most prevalently used methods reported in the literature
for Gender Classification.</p>
      <p>Gender classifications with SVM perform on the raw
image pixels along with different well known feature
extraction methods, namely DTCWT, PCA, and HOG. Training is
done separately on two widely differing datasets consisting
of gender labeled human faces: Color FERET, a set of
images taken under similar conditions with good image quality,
and Adience, a set of labeled unfiltered images intended to
be especially challenging for modern machine learning
algorithms. Adience was designed to present all variations in
appearance, noise, pose, and lighting, that can be expected
of images taken without careful preparation or posing.
(Eidinger, Enbar, and Hassner 2014)</p>
      <p>We use the following steps in conducting our
experiments:
• Uniformly shuffle the order of images.
• Use 70% as training set. 30% as testing set.
• Train with training set.
• Record correct classification rate on testing set.</p>
      <p>Steps 1-4 are repeated 10 times for each experiment using
freshly initialized classifiers.</p>
      <p>We report the results of 18 experiments, 16 of which use
SVMs and two of which use CNN.</p>
    </sec>
    <sec id="sec-4">
      <title>SVM classification</title>
      <p>
        Both linear and RBF kernels were used, with each
constituting a separate experiment using the SVC implementation
included as part of scikit-learn
        <xref ref-type="bibr" rid="ref18">(Pedregosa et al. 2011)</xref>
        with
C = 100 parameter set.
      </p>
      <p>In one experiment, raw pixels are fed into the SVM. Other
experiments used the following feature extraction methods:
PCA, HOG, and DTCWT. Feature extraction was applied
to images uniformly without using face finding software to
isolate and align the face.</p>
    </sec>
    <sec id="sec-5">
      <title>Histogram of Oriented Gradients</title>
      <p>
        HOG descriptors, combined with SVM classifiers, can be
used as a global feature extraction mechanism (Torrione et
al. 2014), while HOG descriptors can be used on locations
indicated by landmark-finding software in areas such as
facial expression classification
        <xref ref-type="bibr" rid="ref7">(De´niz et al. 2011)</xref>
        .
      </p>
      <p>
        One application of HOG descriptors is the automatic
detection of pedestrians, which is made easier in part because
of their predominantly upright pose
        <xref ref-type="bibr" rid="ref6">(Dalal and Triggs 2005)</xref>
        .
We use the standard HOG implementation from the
scikitimage library (van der Walt et al. 2014).
      </p>
      <p>For every image in the Adience and FERET databases,
HOG descriptors were uniformly calculated. 9 orientation
bins were used, and each histogram was calculated based on
gradient orientations in the 7x7 pixel non-overlapping cells.
Normalization was done within each cell (i.e., 1 x 1). The
result was fed into a SVM (SVC class from scikit-learn).</p>
      <p>Training and testing on both Adience and FERET was
performed separately. 30% of images in each database were
used for testing, and the rest for training. For each database,
after reading the data into arrays, the arrays were shuffled
and then the testing and training set were separated.
Training and testing were repeated 10 times with freshly shuffled
data.</p>
    </sec>
    <sec id="sec-6">
      <title>Principal Component Analysis</title>
      <p>PCA is a statistical method for finding correlations between
features in data. When used on images of faces the resulting
images are often referred to as Eigenfaces. PCA is used for
reducing dimensionality of data by eliminating non-essential
information from the dataset and is frequently used in both
image processing and machine learning.</p>
      <p>
        To create the Eigenfaces we used the RandomizedPCA
tool within scikit-learn, which is based on work by
        <xref ref-type="bibr" rid="ref11 ref17">(Halko,
Martinsson, and Tropp 2011)</xref>
        and
        <xref ref-type="bibr" rid="ref11 ref17">(Martinsson, Rokhlin, and
Tygert 2011)</xref>
        . The resulting Eigenfaces were then used in a
linear and RBF SVM.
      </p>
    </sec>
    <sec id="sec-7">
      <title>Convolutional Neural Network</title>
      <p>For the learning stage we used a convolutional neural
network with 3 hidden convolutional layers and one softmax
layer. The training was done using a GTX Titan X GPU
using the Theano based library Pylearn2 and CUDNN
libraries. Stochastic gradient descent was used as the training
algorithm with a momentum of 0.95, found by trial and error.
Learning rates under 0.001 did not show any improvement.
Increasing the learning rate above around 0.005 results in
decreased classification accuracy.</p>
      <p>A general outline of the structure of our CNN is:
• Hidden layer 1: A Rectified Linear Convolutional Layer
using a kernel shape of 4x4, a pool shape of 2x2, a pool
stride of 2x2 and 128 output channels. Initial weights are
randomly selected with a range of 0.5.
• Hidden layer 2: A Rectified Linear Convolutional Layer
using a kernel shape of 4x4, a pool shape of 2x2, a pool
stride of 2x2 and 256 output channels. Initial weights are
randomly selected with a range of 0.5.
• Hidden layer 3: A Rectified Linear Convolutional Layer
using a kernel shape of 3x3, a pool shape of 2x2, a pool
stride of 2x2 and 512 output channels. Initial weights are
randomly selected with a range of 0.5.
• Softmax layer: Initial weights randomly set between 0 and
0.5. Output is the class (male or female).</p>
    </sec>
    <sec id="sec-8">
      <title>Experimental results</title>
      <p>Tables 1 and 2 summarize the classification accuracy of each
approach on each data-set after random shuffling and
separation into 70% training and 30% testing sets. For each
method the grayscale pixels were used as the features,
either directly to the classifier, or to the filter mentioned. For
example, HOG+SVM[RBF] indicates that we use the pixels
as input to a HOG filter, the output of which is used as the
input to a SVM with an RBF kernel.</p>
      <p>DTCWT was both the second best method (after CNN)
and the very worst method we examined; its performance
has the greatest degree of variability depending on the
dataset. It performs very well when objects are consistent
in location and scale. CNN outperformed all methods. Even
the worst CNN experiment on the most difficult dataset
performed better than the best of any other method on the
easiest dataset. This is not a surprising outcome. We wanted
to see if HOG alone was sufficient to increase classification
accuracy as a filter. We found that HOG filters with SVM,
without the usual additional models, provide no benefit on
their own over raw pixel values for this experimental setup.</p>
      <sec id="sec-8-1">
        <title>Method</title>
        <p>CNN
PCA+SVM[RBF]
SVM[RBF]
HOG + SVM[RBF]
HOG+SVM[linear]
PCA+ SVM[linear]
SVM[linear]
DTCWT on SVM[RBF]
DTCWT on SVM[linear]</p>
      </sec>
      <sec id="sec-8-2">
        <title>Method</title>
        <p>CNN
DTCWT on SVM[RBF]
PCA+SVM[RBF]
SVM[RBF]
HOG+SVM[RBF]
HOG+SVM[linear]
DTCWT on SVM[linear]
PCA+SVM[linear]
SVM[linear]</p>
        <p>Mean
96.1%
77.4%
77.3%
75.8%
75%
72 %
70.2%
68.5%
59%</p>
        <p>PCA ties with DTCWT on the best performance on FERET
but performs better than DTCWT on Adiance. As expected
RBF methods performed better than linear SVM classifiers,
however unexpectedly this did not hold true for Adiance,
where differences in filters were enough to cancel out the
effect of RBF in some cases. Every time we used a filter
on FERET RBF was better than linear with filters. This did
not hold for Adience. None of the filters worked particularly
well on Adience, with only PCA slightly outperforming raw
pixels for the RBF classifier.</p>
        <p>On the FERET dataset DTCWT is better (90% vs 86%).
On Adience, it is worse ( 6˜7% vs 77%). This would lend
support to the idea that DTCWT seems to work better (in
theory) on images that are more similar to FERET (uniform
lighting, no complex backgrounds, no extreme warping,
pixelation, or blurring ).</p>
        <p>Using an initial momentum of 0.95 tended to promote fast
convergence without getting stuck in local minimum. We
use a momentum of 0.95 and a learning rate of 0.001.</p>
        <p>
          Using this setup we have achieved an average valid
classification rate of 98% on FERET and 96% on Adience which
is better than the previous highest reported results according
to
          <xref ref-type="bibr" rid="ref16">(Levi and Hassner 2015)</xref>
          on Adience, but we do not
recommend direct comparison of our results with theirs because
of different experimental protocols used.
        </p>
        <p>One of our aims is to investigate the use of the dual tree
complex wavelet transform (DTCWT) on the face feature
classification task. Several recent papers report success in
using DTCWT in gender recognition from frontal face
images citing the benefits of partial rotation invariance. It is
somewhat unclear how to best use this for “In the wild”
images.</p>
      </sec>
    </sec>
    <sec id="sec-9">
      <title>Conclusions</title>
      <p>Much of the previous work on automatic gender
classification use differing datasets and experimental protocols that
can make direct comparisons between reported results
misleading. We have compared nine different machine learning
methods used in gender recognition on two benchmarks,
using identical research methodology to allow a direct
comparison between the efficacies of the different classifiers and
feature extraction methods. In addition to providing updated
information on the effectiveness of these algorithms, we
provide directly comparable results.</p>
      <p>The aim of our study was to explore gender
classification using recent learning algorithms. We carried out
experiments on several state-of-the-art gender classification
methods. We compared the accuracy of these methods on two
very different data sets (“In the wild” verses posed images).</p>
      <p>
        To the extent of our knowledge, this is the first use of
DTCWT on a large 15, 000 database of “in the wild”
images, specifically addressing gender classification. We have
achieved an average accuracy of 98% (FERET) and 96%
(Adience), which is better than the previous highest reported
results (according to
        <xref ref-type="bibr" rid="ref16">(Levi and Hassner 2015)</xref>
        ) on Adience
using a CNN.
      </p>
      <p>The DTCWT seems to work better (⇡ 90%) on images
that are more similar to FERET (uniform lighting, no
complex backgrounds, no extreme warping, pixelation, or
blurring).</p>
      <p>The Adience and FERET data sets are relatively large and
this may explain why the CNN method generally
outperforms other methods: it is known that deep learning
performs well when large training sets are being used. It is
interesting to determine in this particular application what
“large” actually is.
for eigen-feature selection. In Neural Networks, 2002.
IJCNN’02. Proceedings of the 2002 International Joint
Conference on, volume 3, 2433–2438. IEEE.</p>
      <p>Torrione, P. A.; Morton, K. D.; Sakaguchi, R.; and Collins,
L. M. 2014. Histograms of oriented gradients for landmine
detection in ground-penetrating radar data. Geoscience and
Remote Sensing, IEEE Transactions on 52(3):1539–1550.</p>
      <p>Verma, A., and Vig, L. 2014. Using convolutional neural
networks to discover cogntively validated features for
gender classification. In Soft Computing and Machine
Intelligence (ISCMI), 2014 International Conference on, 33–37.
IEEE.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          2011.
          <article-title>Revisiting linear discriminant techniques in gender recognition</article-title>
          .
          <source>Pattern Analysis and Machine Intelligence</source>
          , IEEE Transactions on
          <volume>33</volume>
          (
          <issue>4</issue>
          ):
          <fpage>858</fpage>
          -
          <lpage>864</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Belhumeur</surname>
            ,
            <given-names>P. N.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Hespanha</surname>
            ,
            <given-names>J. P.</given-names>
          </string-name>
          ; and
          <string-name>
            <surname>Kriegman</surname>
            ,
            <given-names>D. J.</given-names>
          </string-name>
          <year>1997</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <article-title>Eigenfaces vs. fisherfaces: Recognition using class specific linear projection</article-title>
          .
          <source>Pattern Analysis and Machine Intelligence</source>
          , IEEE Transactions on
          <volume>19</volume>
          (
          <issue>7</issue>
          ):
          <fpage>711</fpage>
          -
          <lpage>720</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Carcagn</surname>
            `ı,
            <given-names>P.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Del Coco</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Leo</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ; and
          <string-name>
            <surname>Distante</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <article-title>Facial expression recognition and histograms of oriented gradients: a comprehensive study</article-title>
          .
          <source>SpringerPlus</source>
          <volume>4</volume>
          (
          <issue>1</issue>
          ):
          <fpage>1</fpage>
          -
          <lpage>25</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Dalal</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Triggs</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          <year>2005</year>
          .
          <article-title>Histograms of oriented gradients for human detection</article-title>
          .
          <source>In Computer Vision and Pattern Recognition</source>
          ,
          <year>2005</year>
          .
          <article-title>CVPR 2005</article-title>
          . IEEE Computer Society Conference on, volume
          <volume>1</volume>
          ,
          <fpage>886</fpage>
          -
          <lpage>893</lpage>
          . IEEE.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>De´niz</surname>
          </string-name>
          , O.;
          <string-name>
            <surname>Bueno</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Salido</surname>
          </string-name>
          , J.; and
          <string-name>
            <surname>De la Torre</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          <year>2011</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <source>Pattern Recognition Letters</source>
          <volume>32</volume>
          (
          <issue>12</issue>
          ):
          <fpage>1598</fpage>
          -
          <lpage>1603</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Eidinger</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Enbar</surname>
            , R.; and Hassner,
            <given-names>T.</given-names>
          </string-name>
          <year>2014</year>
          .
          <article-title>Age and gender estimation of unfiltered faces</article-title>
          .
          <source>Information Forensics and Security</source>
          , IEEE Transactions on
          <volume>9</volume>
          (
          <issue>12</issue>
          ):
          <fpage>2170</fpage>
          -
          <lpage>2179</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Gutta</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ; Wechsler, H.; and Phillips,
          <string-name>
            <surname>P. J.</surname>
          </string-name>
          <year>1998</year>
          .
          <article-title>Gender and ethnic classification of face images</article-title>
          .
          <source>In Automatic Face and Gesture Recognition</source>
          ,
          <year>1998</year>
          . Proceedings. Third IEEE International Conference on,
          <fpage>194</fpage>
          -
          <lpage>199</lpage>
          . IEEE.
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Halko</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Martinsson</surname>
          </string-name>
          , P.-G.; and
          <string-name>
            <surname>Tropp</surname>
            ,
            <given-names>J. A.</given-names>
          </string-name>
          <year>2011</year>
          .
          <article-title>Finding structure with randomness: Probabilistic algorithms for constructing approximate matrix decompositions</article-title>
          .
          <source>SIAM review 53</source>
          <volume>(2)</volume>
          :
          <fpage>217</fpage>
          -
          <lpage>288</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>Jain</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Huang</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ; and
          <string-name>
            <surname>Fang</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          <year>2005</year>
          .
          <article-title>Gender identification using frontal facial images</article-title>
          .
          <source>In Multimedia and Expo</source>
          ,
          <year>2005</year>
          .
          <article-title>ICME 2005</article-title>
          . IEEE International Conference on, 4-pp.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>Kalansuriya</surname>
            ,
            <given-names>T. R.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Dharmaratne</surname>
            ,
            <given-names>A. T.</given-names>
          </string-name>
          <year>2014</year>
          .
          <article-title>Neural network based age and gender classification for facial images</article-title>
          .
          <source>ICTer</source>
          <volume>7</volume>
          (
          <issue>2</issue>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Kingsbury</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          <year>2001</year>
          .
          <article-title>Complex wavelets for shift invariant analysis and filtering of signals</article-title>
          .
          <source>Applied and Computational Harmonic Analysis</source>
          <volume>10</volume>
          (
          <issue>3</issue>
          ):
          <fpage>234</fpage>
          -
          <lpage>253</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Laytner</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Ling</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ; and
          <string-name>
            <surname>Xiao</surname>
            ,
            <given-names>Q.</given-names>
          </string-name>
          <year>2014</year>
          .
          <article-title>Robust face detection from still images</article-title>
          .
          <source>In Computational Intelligence in Biometrics and Identity Management (CIBIM)</source>
          ,
          <source>2014 IEEE Symposium on</source>
          ,
          <fpage>76</fpage>
          -
          <lpage>80</lpage>
          . IEEE.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>Levi</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Hassner</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          <year>2015</year>
          .
          <article-title>Age and gender classification using convolutional neural networks</article-title>
          .
          <source>In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops</source>
          ,
          <fpage>34</fpage>
          -
          <lpage>42</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Martinsson</surname>
          </string-name>
          , P.-G.;
          <string-name>
            <surname>Rokhlin</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ; and Tygert,
          <string-name>
            <surname>M.</surname>
          </string-name>
          <year>2011</year>
          .
          <article-title>A randomized algorithm for the decomposition of matrices</article-title>
          .
          <source>Applied and Computational Harmonic Analysis</source>
          <volume>30</volume>
          (
          <issue>1</issue>
          ):
          <fpage>47</fpage>
          -
          <lpage>68</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <surname>Pedregosa</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Varoquaux</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Gramfort</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Michel</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Thirion</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Grisel</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Blondel</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Prettenhofer</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ; Weiss, R.;
          <string-name>
            <surname>Dubourg</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Vanderplas</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Passos</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Cournapeau</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ; Brucher,
          <string-name>
            <given-names>M.</given-names>
            ;
            <surname>Perrot</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            ; and
            <surname>Duchesnay</surname>
          </string-name>
          ,
          <string-name>
            <surname>E.</surname>
          </string-name>
          <year>2011</year>
          .
          <article-title>Scikitlearn: Machine learning in Python</article-title>
          .
          <source>Journal of Machine Learning Research</source>
          <volume>12</volume>
          :
          <fpage>2825</fpage>
          -
          <lpage>2830</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Perez</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Tapia</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ; Este´vez, P.; and
          <string-name>
            <surname>Held</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          <year>2012</year>
          .
          <article-title>Gender classification from face images using mutual information and feature fusion</article-title>
          .
          <source>International Journal of Optomechatronics</source>
          <volume>6</volume>
          (
          <issue>1</issue>
          ):
          <fpage>92</fpage>
          -
          <lpage>119</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <surname>Phillips</surname>
            ,
            <given-names>P. J.</given-names>
          </string-name>
          ; Moon, H.; Rizvi,
          <string-name>
            <surname>S. A.</surname>
          </string-name>
          ; and Rauss,
          <string-name>
            <surname>P. J.</surname>
          </string-name>
          <year>2000</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <article-title>The feret evaluation methodology for face-recognition algorithms</article-title>
          .
          <source>Pattern Analysis and Machine Intelligence</source>
          , IEEE Transactions on
          <volume>22</volume>
          (
          <issue>10</issue>
          ):
          <fpage>1090</fpage>
          -
          <lpage>1104</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          2014.
          <article-title>Adaptive multi-stream score fusion for illumination invariant face recognition</article-title>
          .
          <source>In Computational Intelligence in Biometrics and Identity Management (CIBIM)</source>
          ,
          <source>2014 IEEE Symposium on</source>
          ,
          <fpage>94</fpage>
          -
          <lpage>101</lpage>
          . IEEE.
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <surname>Sun</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Yuan</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Bebis</surname>
          </string-name>
          , G.; and
          <string-name>
            <surname>Loui</surname>
            ,
            <given-names>S. J.</given-names>
          </string-name>
          <year>2002</year>
          .
          <article-title>Neuralnetwork-based gender classification using genetic search Wu</article-title>
          , Y.;
          <string-name>
            <surname>Zhuang</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Long</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ;
          <string-name>
            <surname>Lin</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ; and
          <string-name>
            <surname>Xu</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          <year>2015</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <article-title>Human gender classification: A review</article-title>
          .
          <source>arXiv preprint arXiv:1507</source>
          .
          <fpage>05122</fpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>