<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Analogy Making and Logical Inference on Images using Cellular Automata based Hyperdimensional Computing</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Ozgur Yilmaz</string-name>
          <email>ozyilmaz@turgutozal.edu.tr</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Department of Computer Engineering Turgut Ozal University Ankara</institution>
          <country country="TR">Turkey</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2004</year>
      </pub-date>
      <abstract>
        <p>In this paper, we introduce a framework of reservoir computing that is capable of both connectionist machine intelligence and symbolic computation. Cellular automaton is used as the reservoir of dynamical systems. A cellular automaton is a very sparsely connected network with logical nodes and nonlinear/logical connection functions, hence the proposed system corresponds to a binary valued and nonlinear neuro-symbolic architecture. Input is randomly projected onto the initial conditions of automaton cells and nonlinear computation is performed on the input via application of a rule in the automaton for a period of time. The evolution of the automaton creates a space-time volume of the automaton state space, and it is used as the reservoir. In addition to being used as the feature representation for pattern recognition, binary reservoir vectors can be combined using Boolean operations as in hyperdimensional computing, paving a direct way symbolic processing. To demonstrate the capability of the proposed system, we make analogies directly on image data by asking 'What is the Automobile of Air'?, and make logical inference using rules by asking 'Which object is the largest?' Web: ozguryilmazresearch.net 1The literature review is narrowed down in this paper due to space considerations. Please visit our published papers to get a wider view of our architecture among existing reservoir and hyperdimensional computing approaches.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>We have introduced a holistic intelligence framework capable of simultaneous pattern recognition
Yilmaz (2015b) and symbolic computation Yilmaz (2015a,c). Cellular automaton is the main
computational block that holds a distributed representation of high order input attribute statistics as in
neural architectures (Figure 1 b). The proposed architecture is a cross fertilization of cellular
automata, reservoir computing and hyperdimensional computing frameworks (Figure 1 a). The cellular
automata (CA) computation can be viewed as a feedforward network with logical nodes and
connections, as shown in Figure 1 c. In this paper, we analyze the symbolic computation capability of
the system on making analogies and rule based logical inferences, directly on the image data. The
results show that (Figure 2), binary vector representation of images derived through CA evolution
provide very precise analogies and accurate rule based inference, even though very small number of
examples are provided. In the next subsection we review cellular automata 1, then introduce relevant
neuro-symbolic computation studies. Finally, we state our contribution.</p>
      <p>Input</p>
      <p>Win</p>
      <p>CA Rule
Reservoir
Computing</p>
      <p>Hyperdimensional</p>
      <p>Computing
time
Cellular Automata</p>
      <p>Reservoir</p>
      <p>Wout</p>
      <p>Output</p>
      <p>c
Cellular automaton is a discrete computational model consisting of a regular grid of cells, each in one
of a finite number of states (Figure 1 c). The state of an individual cell evolves in time according to
a fixed rule, depending on the current state and the state of its neighbors. The information presented
as the initial states of a grid of cells is processed in the state transitions of cellular automaton and
computation is typically very local. Essentially, a cellular automaton is a very sparsely connected
network with logical nodes and nonlinear/logical connection functions (Figure 1 c). Some of the
cellular automata rules are proven to be computationally universal, capable of simulating a Turing
machine (Cook, 2004).</p>
      <p>
        The rules of cellular automata are classified according to their behavior: attractor, oscillating,
chaotic, and edge of chaos (Wolfram, 2002). Turing complete rules are generally associated with the
last class (rule 110, Conway game of life). Lyapunov exponent of a cellular automaton can be
computed and it is shown to be a good indicator of the computational power of the automata
        <xref ref-type="bibr" rid="ref2">(Baetens &amp;
De Baets, 2010)</xref>
        . A spectrum of Lyapunov exponent values can be achieved using different cellular
automata rules. Therefore, a dynamical system with specific memory capacity (i.e. Lyapunov
exponent value) can be constructed by using a corresponding cellular automaton. The time evolution of
the cellular automata has very rich computational representation
        <xref ref-type="bibr" rid="ref18">Mitchell et al. (1996)</xref>
        , especially for
the edge of chaos dynamics. The proposed algorithm in this paper exploits the entire time evolution
of the CA and uses the states as the reservoir
        <xref ref-type="bibr" rid="ref13">LukosˇEvicˇIus &amp; Jaeger (2009)</xref>
        ;
        <xref ref-type="bibr" rid="ref14">Maass et al. (2002)</xref>
        of
nonlinear computation.
1.2
      </p>
      <sec id="sec-1-1">
        <title>Symbolic Computation on Neural Representations</title>
        <p>
          Uniting the expressive power of mathematical logic and pattern recognition capability of distributed
representations (eg. neural networks) has been an open question for decades although several
successful theories have been proposed
          <xref ref-type="bibr" rid="ref1 ref15 ref16 ref21 ref3 ref8">(Garcez et al., 2012; Bader et al., 2008; Marcus, 2003;
Miikkulainen et al., 2006; Besold et al., 2014; Pollack, 1990)</xref>
          . The difficulty arises due to the very
different mathematical nature of logical reasoning and dynamical systems theory. Along with many
other researchers, we conjecture that combining connectionist and symbolic processing requires
commonalizing the representation of data and knowledge.
        </p>
        <p>
          Along the same vein,
          <xref ref-type="bibr" rid="ref9">(Kanerva, 2009)</xref>
          introduced hyperdimensional computing that utilizes
highdimensional random binary vectors for representing objects, predicates and rules for symbolic
manipulation and inference. The general family of the methods is called ’reduced representations’ or
’vector symbolic architectures’, and detailed introductions can be found in
          <xref ref-type="bibr" rid="ref12 ref20">(Plate, 2003; Levy &amp;
Gayler, 2008)</xref>
          . In this approach, high dimensionality and randomness enable binding and grouping
operations that are essential for one shot learning, analogy-making, hierarchical concept building
and rule based logical inference. Most recently
          <xref ref-type="bibr" rid="ref7">(Gallant &amp; Okaywe, 2013)</xref>
          introduced random
matrices to this context and extended the binding and quoting operations. The two basic mathematical
tools of reduced representations are vector addition and XOR.
        </p>
        <p>In this paper, we borrow these tools of hyperdimensional computing framework, and build a
semantically more meaningful representation by removing the randomness and replacing it with the
cellular automata computation. This provides not only a more expressive symbolic computation
architecture, but also enables pattern recognition capabilities otherwise not possible with random
vectors.
1.3</p>
      </sec>
      <sec id="sec-1-2">
        <title>Contributions</title>
        <p>We provide a low computational complexity method Yilmaz (2015b) for recurrent computation
using cellular automata based hyperdimensional computing. It is shown that the framework has a great
potential for symbolic processing such that the cellular automata feature space can directly be
combined by Boolean operations as in hyperdimensional computing, hence they can represent concepts
and form a hierarchy of semantic interpretations. We demonstrate this capability by making
analogies directly on images and infer relationships using logical rules. In the next section, we give the
details of the algorithm, and then provide results experiments that demonstrate our contributions.
2</p>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>Cellular Automata Feature Expansion</title>
      <p>In our reservoir computing method, data are passed on a cellular automaton instead of an echo
state network and the nonlinear dynamics of cellular automaton provide the necessary projection
of the input data onto an expressive and discriminative space. Compared to classical neuron-based
reservoir computing, the reservoir design is trivial: cellular automaton rule selection. Utilization of
edge of chaos automaton rules ensures Turing-complete computation in the reservoir, which is not
guaranteed in classical reservoir computing approaches.</p>
      <p>The reservoir computing system receives the input data. First, the encoding stage translates the
input into the initial states of a 1D elementary cellular automaton. For binary input data, each
feature dimension can randomly be mapped onto the cells of the cellular automaton. For this type
of mapping, the size of the CA should follow the input data’s feature dimension. After encoding,
suppose that the cellular automaton is initialized with the vector A0P1 , in which P1 corresponds
to a random permutation of raw input data. Then, cellular automata evolution is computed using a
prespecified rule, Z (1D elementary CA rule), for a fixed period of iterations (I):
A1P1 = Z(A0P1 );
A2P1 = Z(A1P1 );
:::</p>
      <p>AI P1 = Z(AI 1P1 ):
The evolution of the cellular automaton is recorded such that, at each time step a snapshot of the
whole states in the cellular automaton is vectorized and concatenated. Therfore, we concatenate the
evolution of cellular automata to obtain a reservoir for a single permutation:</p>
      <p>AP1 = [A0P1 ; A1P1 ; A2P1 ; :::AI P1 ]
It is experimentally observed that multiple random permutation mappings significantly improve
accuracy. There are R number of different random mappings, i.e., separate CA reservoirs, and they
are combined into a large reservoir feature vector:</p>
      <p>AR = [AP1 ; AP2 ; AP3 ; :::APR ]:
The computation in CA takes place when cell activities due to nonzero initial values (i.e., input)
mix and interact. Both prolonged evolution duration (large I ) and existence of different random
mappings (large R) increase the probability of long-range interactions, hence improve computational
power and enhance representation.
3</p>
    </sec>
    <sec id="sec-3">
      <title>Symbolic Processing and Non-Random Hyperdimensional Computing</title>
      <p>Hyperdimensional computing uses random very large-sized binary vectors to represent objects,
concepts, and predicates. Then, appropriate binding and grouping operations are used to manipulate the
vectors for hierarchical concept building, analogy-making, learning from a single example, etc., that
are hallmarks of symbolic computation. The large size of the vector provides a vast space of random
vectors, two of which are always nearly orthogonal. Yet, the code is robust against a distortion in
the vector due to noise or imperfection in storage because after distortion it will still stay closer to
the original vector than the other random vectors.</p>
      <p>The grouping operation is normalized vector summation, and it enables forming sets of
objects/concepts. Suppose we want to group two binary vectors, V1 and V2. We compute their
elementwise sums, and the resultant vector contains 0, 1 and 2 entries. We normalize the vector by accepting
the 0 entries as they are, transforming 2 entries into 1. Note that, these are consistent within the two
initial vectors. Then, the inconsistent entries are randomly decided: 1 entries as transformed into 0
or 1. Many vectors can be combined iteratively or in a batch to form a grouped representation of the
bundle. The resultant vector is similar to all the elements of the vector due to the fact that consistent
entries are untouched. The elements of the set can be recovered from the reduced representation by
probing with the closest item in the memory, and consecutive subtraction. Grouping is essential for
defining ’a part of’, ’contains’ relationships. + symbol will be used for normalized summation in
the following arguments.</p>
      <p>
        There are two binding operations: bitwise XOR (circled plus symbol, ) and permutation 2.
Binding operation maps (randomizes) the vector to a completely different space, while preserving the
distances between two vectors. As stated in
        <xref ref-type="bibr" rid="ref9">Kanerva (2009)</xref>
        , ”...when a set of points is mapped
by multiplying with the same vector, the distances are maintained, it is like moving a constellation
of points bodily into a different (and indifferent) part of the space while maintaining the relations
(distances) between them. Such mappings could play a role in high-level cognitive functions such as
analogy and the grammatical use of language where the relations between objects is more important
than the objects themselves.
      </p>
      <p>A few representative examples to demonstrate the expressive power of hyperdimensional computing:
1. We can represent pairs of objects via multiplication. OA;B = A
object vectors.</p>
      <sec id="sec-3-1">
        <title>B where A and B are two</title>
        <p>
          2. A triplet is a relationship between two objects, defined by a predicate. This can similarly be
formed by TA;P;B = A P B. These types of triplet relationships are very successfully utilized
for information extraction in large knowledge bases
          <xref ref-type="bibr" rid="ref6">Dong et al. (2014)</xref>
          .
3. A composite object can be built by binding with attribute representation and summation. For a
composite object C,
        </p>
        <p>C = X</p>
        <p>A1 + Y</p>
        <p>A2 + Z</p>
        <p>
          A3;
where A1, A2 and A3 are vectors for attributes and X , Y and Z are the values of the attributes for a
specific composite object.
4. A value of an attribute for a composite object can be substituted by multiplication. Suppose we
have assignment X A1, then we can substitute A1 with B1 by, (X A1) (A1 B1) = X B1.
It is equivalent to say that A1 and B1 are analogous. This property is essential for analogy making.
2Please see
          <xref ref-type="bibr" rid="ref9">Kanerva (2009)</xref>
          for the details of permutation operation, as a way of doing multiplication.
5. We can define rules of inference by binding and summation operations. Suppose we have a rule
stating that ”If x is the mother of y and y is the father of z, then x is the grandmother of z” 3. Define
atomic relationships:
then the rule is,
Given the knowledge base, ”Anna is the mother of Bill” and ”Bill is the father of Cid”, we can infer
grandmother relationship by applying the rule Rxyz:
        </p>
        <p>Mxy = M1
Fyz = F1
Gxz = G1</p>
        <p>X + M2</p>
        <p>Y;
Y + M2
X + G2</p>
        <p>Z;</p>
        <p>Z;
Rxyz = Gxz</p>
        <p>(Mxy + Fyz):
Mab = M1
Fbc = F1</p>
        <p>0
Gac = Rxyz</p>
        <p>A + M2
B + M2</p>
        <p>B;</p>
        <p>C;
(Mab + Fbc);
where vector G0ac is expected to be very similar to Gac, which says ”Anna is the grandmother of
Cid”. Please note that the if-then rules represented by hyperdimensional computing can only be
if-and-only-if logical statements because operations used to represent the rules are symmetric.
Without losing the expressive power of classical hyperdimensional computing, we are introducing
cellular automata to the framework. In our approach, we use binary cellular automata reservoir
vector as the representation of objects and predicates instead of random vectors to be used for
symbolic computation. There are two major advantages of this approach over random binary vector
generation:
1. Reservoir vector enables connectionist pattern recognition and statistical machine learning (as
demonstrated in Yilmaz (2015b)) while random vectors are mainly tailored for symbolic
computation.
2. The composition and modification of objects can be achieved in a semantically more meaningful
way. The semantic similarity of the two data instances can be preserved in the reservoir
hyperdimensional vector representation, but there is no straightforward mechanism for this in the classical
hyperdimensional computing framework.
4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Experiments on Analogy Making</title>
      <p>
        In order to demonstrate the power of enabled logical operation, we will use analogy making.
Analogy making is crucial for generalization of what is already learned. We tested the capability of our
symbolic system using images. The example given here follows ”What is the Dollar of Mexico?”
in
        <xref ref-type="bibr" rid="ref9">Kanerva (2009)</xref>
        . However in the original example, sensory data (i.e. image) is not considered
because there is no straightforward way to introduce sensory data into the hyperdimensional
computing framework. The benefit of using non-random binary vectors is obvious in this context.
We have previously shown that binarization of the hidden layer activities of a feedforward network
is not very detrimental for classification purposes
        <xref ref-type="bibr" rid="ref22 ref23 ref24 ref25">Yilmaz et al. (2015)</xref>
        . For an image, the binary
representation of the first hidden layer activities holds an indicator for the existence of Gabor like
corner and edge features. In order to test the symbolic computation performance of CA features on
binarized hidden layer activities, we use CIFAR 10 dataset
        <xref ref-type="bibr" rid="ref11">Krizhevsky &amp; Hinton (2009)</xref>
        . We used
the first 500 training/test images and obtained single layer hidden neuron representation using the
algorithm in
        <xref ref-type="bibr" rid="ref4">Coates et al. (2011)</xref>
        (200 number of different receptive fields, receptive fields size of 6
pixels). The neural activities are binarized according to a threshold and, on average, 22 percent of
the neurons fired with the selected threshold. After binarization of neural activities, CA features can
be computed on the binary representation as explained in section 2. We formed a separate concept
vector for each class (total 10 classes, 50 examples for each class) using binary neural representation
of CIFAR training data and vector addition defined in Snaider (2012). These are the basis class
concepts extracted from the visual database.
      </p>
      <p>
        3The example is adapted from
        <xref ref-type="bibr" rid="ref9">Kanerva (2009)</xref>
        .
      </p>
      <sec id="sec-4-1">
        <title>We formed two new concepts called Land and Air:</title>
        <p>Land = Animal</p>
        <p>Horse + V ehicle</p>
        <p>Automobile;
Air = Animal</p>
        <p>Bird + V ehicle</p>
        <p>Airplane:
In these two concepts, CA features of Horse and Bird images are used to bind with the Animal
filler, and CA features of Automobile and Airplane images are used to bind with the Vehicle filler
4. Animal and Vehicle fields are represented by two random vectors 5, those with the same size as
the CA features. Multiplication is performed by xor ( ) operation and vector summation is again
identical to Snaider (2012). The products, Land and Air are also CA feature vectors, and they
represent the merged concept of observed animals and vehicles in Land and Air respectively. We
can ask the analogical question ”What is the Automobile of Air?”, AoA in short. The answer can
simply be given by this equality (inference):</p>
        <p>AoA = Automobile</p>
        <p>Land</p>
        <p>Air:
AoA is a CA feature vector and expected to be very similar to Airplane concept vector. We tested the
analogical accuracy using unseen Automobile test images (50 in total), computing their CA feature
vectors followed by AoA inference, then finding the closest concept class vector to AoA vector
(max inner product). It is expected to be the Airplane class. The result of this experiment is given
in Figure 2 a 6 for various R and I combinations. The multiplication of the two defines the amount
of feature expansion due to cellular automata state space. The analogy on CA features is 98 percent
accurate (for both R and I equals 128), whereas if the binary hidden layer activity is used instead
of CA features (corresponds to R and I equal to 1), the analogy is only 21 percent accurate. This
result clearly demonstrates the benefit of CA feature expansion for symbolic computation.
The devised analogy implicitly assumes that Automobile concept is already encoded in the concept
of Land. What if we ask ”What is the Truck of Air?”? Even though Truck images are not used in
building the Land concept, due to the similarity of Truck and Automobile concepts, we might still
get good analogies. The results on this second order analogy is contrasted in Table 1. Automobile
and Horse (i.e. ”What is the Horse of Air?”, the answer should be Bird.) are first order analogies
and they result in comparably superior performance as expected, but second order analogy is much
higher than chance level (i.e., 10 percent).</p>
        <p>Please note that these analogies are performed strictly on the sensory data, i.e., images. Given an
image, the system is able to retrieve a set of relevant images that are linked through a logical statement.</p>
        <p>4There are 50 training images for each class. CA rule 110 is used for evolution. And mean of 20 Monte
Carlo simulations is given to account for randomness in experiments
5Also 22 percent non-zero elements
6These are extended results for our previous publication Yilmaz (2015a). We were unable to test for large
R and I values due to hardware limitations.</p>
        <sec id="sec-4-1-1">
          <title>Automobile</title>
          <p>79</p>
        </sec>
        <sec id="sec-4-1-2">
          <title>Horse</title>
          <p>68</p>
        </sec>
      </sec>
      <sec id="sec-4-2">
        <title>Truck 52</title>
        <p>A very small number of training data is used, yet we can infer conceptual relationships between
images surprisingly accurately. However, again it should be emphasized that this is only a single
experiment with a very limited analogical scope, and more experiments are needed to understand
the limits of the proposed architecture.</p>
        <p>
          It is possible to build much more complicated concepts using hierarchies. For example, Land and
Air are types of environments and can be used as fillers in Environment field. Ontologies are helpful
to narrow down the set of required concepts for attaining a satisfactory description of the world.
Other modalities such as text data are also of great interest, (see
          <xref ref-type="bibr" rid="ref17">Mikolov et al. (2013)</xref>
          ;
          <xref ref-type="bibr" rid="ref19">Pennington
et al. (2014)</xref>
          for state-of-the-art studies), as well as information fusion on multiple modalities (e.g.,
Image and text).
5
        </p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Experiments on Rule Based Inference</title>
      <p>In order to test proposed architecture’s capability for logical inference, we define a rule and form a
knowledge base on image data. Then we make an inference on the knowledge base by applying the
rule. The inference may or may not be right, hence the symbolic system is not completely sound. 7
For demonstration of logical inference on our system, we define size relationships among different
objects using the following rule set.</p>
      <p>Object in image X is larger than object in image Y :</p>
      <sec id="sec-5-1">
        <title>Object in image X is smaller than object in image Z:</title>
      </sec>
      <sec id="sec-5-2">
        <title>And finally, we state the largest object is in image Z:</title>
        <p>Lxy = L1</p>
        <p>X + L2
Sxz = S1</p>
        <p>X + S2</p>
        <p>Y:</p>
        <p>Z:
Tz = T1</p>
        <p>Z:
Then the rule is stated as ’If object in X is larger than object in Y and smaller than object in Z,
largest object is in Z’. The rule vector is computed as the manipulation of object vectors using
hyperdimensional computing framework:</p>
        <p>Rxyz = Tz
(Lxy + Sxz ):
Our knowledge base is again formed on the images of CIFAR 10 dataset. First, we use Truck,
Automobile, and Airplane images (50 each) to compute X , Y and Z concept vectors (CA rule 110);
then we obtain the rule vector Rxyz as explained above utilizing the concept vectors. In a completely
different set of test image triplet, we make use of single Truck, Automobile, and Airplane images
and compute their vector representation; a, b and c respectively. Knowledge base is created on the
object vectors in three test images:</p>
        <p>Lab = L1
Sac = F1
a + L2
a + M2
b;
c;
as ’object in image a is larger than object in image b, and object in image a is smaller than object
in image c’. Can we infer the largest object? When we apply the rule vector on existing knowledge
base, we get an estimate for the vector representation of the largest object:</p>
        <p>Test = Rxyz
(Lab + Sac):
We compute the Hamming distance of Test to existing object vectors (i.e. a, b and c), then it is
possible to decide on the estimated largest object, i.e. closest vector to Test which should be vector
7The completeness of the system requires a proof and it is a future work.
c. The average accuracy of 50 different test image triplets are shown in Figure 2 b. The chance
level is 33 percent and it is observed that the binary neural representation (i.e., both R and I is equal
to 1) is around 50 percent accurate, whereas cellular automata state space provides a 100 percent
inference accuracy for a relatively small reservoir size. Please note that, similar to analogy making
experiments logical inference is performed directly on image data and we can make object size
inference using a very small number of example images (50).
6</p>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>Discussion</title>
      <p>
        Along with the pattern recognition capabilities of cellular automata based reservoir computing
Yilmaz (2015b), hyperdimensional computing framework enables symbolic processing. Due to the
binary categorical indicator nature of the representation, the rules that make up the knowledge base
and feature representation of the data that make up the statistical model live on the same space,
which is essential for combining connectionist and symbolic capabilities. It is possible to make
analogies, form hierarchies of concepts, and apply logical rules on the reservoir feature vectors 8.
To illustrate the logical query, we have shown the capability of the system to make analogies on
image data. We asked the question ”What is the Automobile of Air?” after building Land and Air
concepts based on the images of Horse, Automobile (Land), Bird and Airplane (Air). The correct
answer is Airplane and the system infers this relationship with 98 percent accuracy, with only 50
training images per class. Additionally we have tested the performance of our architecture on rule
based logical inference on images. We defined an object size related rule on image data, provided a
knowledge base and inferred the largest object strictly using the image features.
Neural network data embeddings (eg.
        <xref ref-type="bibr" rid="ref10">Kiros et al. (2014)</xref>
        ;
        <xref ref-type="bibr" rid="ref17">Mikolov et al. (2013)</xref>
        ) are an alternative
to our approach, in which representation suitable for logical manipulation is learned from the data
using gradient descent. Although these promising approaches are showing state-of-the-art results,
they are bound to suffer from the dilemma of ’no free lunch’ because the representation is
dataspecific. The other extreme is random embeddings adopted in hyperdimensional computing and
reduced vector representation approaches. Although randomness maximizes the orthogonality of
vectors and optimizes the effective usage of the space, it does not allow statistical machine learning
or semantically meaningful modifications on existing vectors. Our approach lies in the middle: it
does not create random vectors, thus can manipulate existing vectors and use machine learning, but
do not learn the representation from the data therefore it is less prone to overfitting as well as to the
dilemma of ’no free lunch’. Moreover, cellular automata reservoir is orders of magnitude faster than
neural network counterparts Yilmaz (2015b).
7
      </p>
    </sec>
    <sec id="sec-7">
      <title>Acknowledgments</title>
      <p>This research is supported by The Scientific and Technological Research Council of Turkey
(TUBI˙TAK) Career Grant, No: 114E554.
Snaider, J. (2012). Integer sparse distributed memory and modular composite representation, .
Wolfram, S. (2002). A new kind of science volume 5. Wolfram media Champaign.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Bader</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hitzler</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          , &amp; Ho¨lldobler,
          <string-name>
            <surname>S.</surname>
          </string-name>
          (
          <year>2008</year>
          ).
          <article-title>Connectionist model generation: A first-order approach</article-title>
          . Neurocomputing,
          <volume>71</volume>
          ,
          <fpage>2420</fpage>
          -
          <lpage>2432</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Baetens</surname>
            ,
            <given-names>J. M.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>De Baets</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>Phenomenological study of irregular cellular automata based on lyapunov exponents and jacobians</article-title>
          .
          <source>Chaos: An Interdisciplinary Journal of Nonlinear Science</source>
          ,
          <volume>20</volume>
          ,
          <fpage>033112</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Besold</surname>
            ,
            <given-names>T. R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Garcez</surname>
            ,
            <given-names>A. d.</given-names>
          </string-name>
          , Ku¨hnberger, K.-U., &amp;
          <string-name>
            <surname>Stewart</surname>
            ,
            <given-names>T. C.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Neural-symbolic networks for cognitive capacities</article-title>
          .
          <source>Biologically Inspired Cognitive Architectures</source>
          , (pp.
          <fpage>iii</fpage>
          -iv).
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Coates</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ng</surname>
            ,
            <given-names>A. Y.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Lee</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>An analysis of single-layer networks in unsupervised feature learning</article-title>
          .
          <source>In International Conference on Artificial Intelligence and Statistics</source>
          (pp.
          <fpage>215</fpage>
          -
          <lpage>223</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <article-title>8Linear CA rules, such as rule 90 allow superposition of initial conditions. This property provides a symbolic system with much more powerful expressive capability Yilmaz (2015a).</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Dong</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gabrilovich</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Heitz</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Horn</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lao</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Murphy</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Strohmann</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sun</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Zhang</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Knowledge vault: A web-scale approach to probabilistic knowledge fusion</article-title>
          .
          <source>In Proceedings of the 20th ACM SIGKDD international conference on Knowledge discovery and data mining</source>
          (pp.
          <fpage>601</fpage>
          -
          <lpage>610</lpage>
          ). ACM.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Gallant</surname>
            ,
            <given-names>S. I.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Okaywe</surname>
            ,
            <given-names>T. W.</given-names>
          </string-name>
          (
          <year>2013</year>
          ).
          <article-title>Representing objects, relations, and sequences</article-title>
          .
          <source>Neural computation</source>
          ,
          <volume>25</volume>
          ,
          <fpage>2038</fpage>
          -
          <lpage>2078</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Garcez</surname>
            ,
            <given-names>A. S. d.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Broda</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Gabbay</surname>
            ,
            <given-names>D. M.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>Neural-symbolic learning systems: foundations and applications</article-title>
          . Springer Science &amp; Business Media.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Kanerva</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>Hyperdimensional computing: An introduction to computing in distributed representation with high-dimensional random vectors</article-title>
          .
          <source>Cognitive Computation</source>
          ,
          <volume>1</volume>
          ,
          <fpage>139</fpage>
          -
          <lpage>159</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Kiros</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Salakhutdinov</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Zemel</surname>
            ,
            <given-names>R. S.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Unifying visual-semantic embeddings with multimodal neural language models</article-title>
          .
          <source>arXiv preprint arXiv:1411</source>
          .
          <fpage>2539</fpage>
          , .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Krizhevsky</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Hinton</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>Learning multiple layers of features from tiny images</article-title>
          . Computer Science Department, University of Toronto, Tech. Rep, .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>Levy</surname>
            ,
            <given-names>S. D.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Gayler</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          (
          <year>2008</year>
          ).
          <article-title>Vector symbolic architectures: A new building material for artificial general intelligence</article-title>
          .
          <source>In Proceedings of the 2008 conference on Artificial General Intelligence 2008: Proceedings of the First AGI Conference</source>
          (pp.
          <fpage>414</fpage>
          -
          <lpage>418</lpage>
          ). IOS Press.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>LukosˇEvicˇIus</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Jaeger</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>Reservoir computing approaches to recurrent neural network training</article-title>
          .
          <source>Computer Science Review</source>
          ,
          <volume>3</volume>
          ,
          <fpage>127</fpage>
          -
          <lpage>149</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Maass</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          , Natschla¨ger,
          <string-name>
            <given-names>T.</given-names>
            , &amp;
            <surname>Markram</surname>
          </string-name>
          ,
          <string-name>
            <surname>H.</surname>
          </string-name>
          (
          <year>2002</year>
          ).
          <article-title>Real-time computing without stable states: A new framework for neural computation based on perturbations</article-title>
          .
          <source>Neural computation</source>
          ,
          <volume>14</volume>
          ,
          <fpage>2531</fpage>
          -
          <lpage>2560</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Marcus</surname>
            ,
            <given-names>G. F.</given-names>
          </string-name>
          (
          <year>2003</year>
          ).
          <article-title>The algebraic mind: Integrating connectionism and cognitive science</article-title>
          . MIT press.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>Miikkulainen</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bednar</surname>
            ,
            <given-names>J. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Choe</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Sirosh</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          (
          <year>2006</year>
          ).
          <article-title>Computational maps in the visual cortex</article-title>
          . Springer Science &amp; Business Media.
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Mikolov</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sutskever</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Corrado</surname>
            ,
            <given-names>G. S.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Dean</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          (
          <year>2013</year>
          ).
          <article-title>Distributed representations of words and phrases and their compositionality</article-title>
          .
          <source>In Advances in Neural Information Processing Systems</source>
          (pp.
          <fpage>3111</fpage>
          -
          <lpage>3119</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <surname>Mitchell</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          et al. (
          <year>1996</year>
          ).
          <article-title>Computation in cellular automata: A selected review</article-title>
          .
          <source>Nonstandard Computation</source>
          , (pp.
          <fpage>95</fpage>
          -
          <lpage>140</lpage>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Pennington</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Socher</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Manning</surname>
            ,
            <given-names>C. D.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Glove: Global vectors for word representation</article-title>
          .
          <source>Proceedings of the Empiricial Methods in Natural Language Processing (EMNLP</source>
          <year>2014</year>
          ),
          <volume>12</volume>
          ,
          <fpage>1532</fpage>
          -
          <lpage>1543</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <surname>Plate</surname>
            ,
            <given-names>T. A.</given-names>
          </string-name>
          (
          <year>2003</year>
          ).
          <article-title>Holographic reduced representation: Distributed representation for cognitive structures</article-title>
          , .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <string-name>
            <surname>Pollack</surname>
            ,
            <given-names>J. B.</given-names>
          </string-name>
          (
          <year>1990</year>
          ).
          <article-title>Recursive distributed representations</article-title>
          .
          <source>Artificial Intelligence</source>
          ,
          <volume>46</volume>
          ,
          <fpage>77</fpage>
          -
          <lpage>105</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <surname>Yilmaz</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          (
          <year>2015a</year>
          ).
          <source>Symbolic Computation using Cellular Automata based Hyperdimensional Computing. Neural Computation</source>
          , .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <surname>Yilmaz</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          (
          <year>2015b</year>
          ).
          <article-title>Machine Learning using Cellular Automata based Feature Expansion and Reservoir Computing</article-title>
          .
          <source>Journal of Cellular Automata</source>
          , .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <string-name>
            <surname>Yilmaz</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          (
          <year>2015c</year>
          ).
          <article-title>Connectionist-Symbolic Machine Intelligence using Cellular Automata based Reservoir-Hyperdimensional Computing</article-title>
          .
          <source>arXiv preprint arXiv:1503</source>
          .
          <fpage>00851</fpage>
          , .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          <string-name>
            <surname>Yilmaz</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ozsarac</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gunay</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Ozkan</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          (
          <year>2015</year>
          ).
          <article-title>Cognitively inspired real-time vision core</article-title>
          .
          <source>Technical Report</source>
          , . Available at http://ozguryilmazresearch.net/ Publications/NeuralNetworkVisionCore_YilmazEtAl2015.pdf.
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>