<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Hebbian Learning Mechanisms Help Explain the Maturation of Multisensory Speech Integration in Children with Autism Spectrum Disorder (ASD) and with Typical Development (TD): a Neurocomputational Analysis.</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Cristiano Cuppini (cristiano.cuppini@unibo.it)</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mauro Ursino (mauro.ursino@unibo.it)</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Lars A. Ross (lars.ross@einstein.yu.edu)</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>John J. Foxe (john.foxe@einstein.yu.edu)</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Elisa Magosso</institution>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Sophie Molholm</institution>
        </aff>
      </contrib-group>
      <fpage>389</fpage>
      <lpage>394</lpage>
      <abstract>
        <p />
      </abstract>
      <kwd-group>
        <kwd>Autism Spectrum Disorder (ASD)</kwd>
        <kwd>Neural Networks</kwd>
        <kwd>Hebbian Learning Rules</kwd>
        <kwd>Multisensory Speech Integration</kwd>
        <kwd>McGurk Effect</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        Cognitive tasks such as communication and speech
comprehension rely on the brain’s ability to exploit and
integrate sensory information of different modalities.
Accordingly, the appropriate development of multisensory
speech integration (MSI) greatly influences a child’s ability to
successfully relate with others. Several experimental findings
have shown that speech intelligibility is affected by
visualizing a speaker’s articulations, and that MSI continues
developing late into childhood. This work aims at developing
a network to analyze the role of the sensory experience during
the early stages of life, as a mechanism responsible for the
maturation of these integrative abilities in teenagers. We
extended a model realized to study multisensory integration in
cortical regions
        <xref ref-type="bibr" rid="ref16 ref25 ref6">(Magosso et al., 2012; Cuppini et al, 2014)</xref>
        by
incorporating a multisensory area known to be involved in
audiovisual speech processing, the superior temporal sulcus
(STS). The model suggests that the maturation of MSI is
primarily due to the maturation of direct connections among
primary unisensory regions. This process was the results of a
training phase during which the network was exposed to
sensory-specific and cross-sensory stimuli, and excitatory
projections among the unisensory regions of the model were
subjected to Hebbian rules of potentiation and depression.
With such a model, we also analyzed the acquisition of adult
MSI abilities in ASD children, and we were able to explain
the delayed maturation as result of a lower level of
multisensory exposures during early phases of life.
      </p>
    </sec>
    <sec id="sec-2">
      <title>Introduction</title>
      <p>
        The brain’s ability to exploit and integrate sensory
information of different modalities is fundamental not just
for simple detection tasks, but also for more demanding
perceptual-cognitive functions, such as those involved in
communication. For example, the intelligibility of speech is
significantly improved when one can see the speaker’s
articulations. Accordingly, the appropriate development of
multisensory speech integration (MSI) greatly affects a
child’s ability to relate with others. Ample experimental
evidence has shown that MSI appears to be highly immature
at birth and that continues to develop late into childhood
(Brandwein et al., 2010). Moreover, children with autism
spectrum disorder (ASD) presenting impaired MSI early in
life, show an amelioration in the adolescent years
        <xref ref-type="bibr" rid="ref3 ref9">(de
BoerSchellekens et al., 2013; Foxe et al., 2015)</xref>
        . These evidences
suggest that there may be delays in the maturation of MSI
for children with ASD that resolve at this point. Multiple
studies have shown multisensory processing deficits in ASD
in the absence of comparable unisensory deficits, suggesting
that they represent impairment of neural processes that have
direct and specific impact on MSI. However, the neural
basis of the impairment remains unknown.
      </p>
      <p>
        A region of particular interest for the maturation of MSI is
the superior temporal sulcus (STS), an association cortex
involved in speech perception (Molholm et al., 2013) that is
also frequently implicated in audiovisual multisensory
processing
        <xref ref-type="bibr" rid="ref4">(Bolognini et al., 2009)</xref>
        . This region must be
considered in the context of its feedforward inputs from
auditory and visual cortices. Converging evidence reveals
that MSI occurs at very early stages of cortical processing
and in sensory cortical regions, although the functional role
of early MSI
        <xref ref-type="bibr" rid="ref19">(at the onset of cortical sensory processing in
some cases; Molholm et al., 2002)</xref>
        remains unknown.
      </p>
      <p>
        Several experimental data pointed out that auditory
speech recognition is relatively mature at 5 to 9 years of
age, approaching adult-like performances
        <xref ref-type="bibr" rid="ref14 ref8">(e.g., Fallon,
Trehub &amp; Schneider, 2000; Kraus, Koch, McGee, Nicol, &amp;
Cunningham, 1999)</xref>
        , at ages where multisensory speech
processing is not
        <xref ref-type="bibr" rid="ref9">(Foxe et al, 2015)</xref>
        .
      </p>
      <p>Such observations suggest a neural model in which the
maturation of MSI in speech perception follows from the
reinforcement of direct “cross-modal” excitatory
connections between auditory and visual speech
representations in unisensory cortices. In this case, it can be
assumed that connections among unisensory areas are
initially relatively ineffective, but that they strengthen as a
consequence of relevant multisensory experiences through a
Hebbian learning mechanism. Thus, multisensory
experiences would affect only the ability of STS elements
to detect multisensory stimuli, via a reciprocal
reinforcement of unisensory activities when both are active,
but it would not produce any additional level of information
to the STS in case of unisensory stimulation.</p>
      <p>
        The aim of the present work is to test the feasibility of this
model, and its consequences by using a computational
model inspired by neurophysiology and based on a previous
model implemented to study cortical multisensory
interaction
        <xref ref-type="bibr" rid="ref16 ref25 ref6">(Magosso et al., 2012; Cuppini et al., 2014)</xref>
        . In
particular, with the model we wish to i) analyze possible
mechanisms underlying the maturation of MSI; ii) test the
model's ability to reproduce different results concerning
speech MSI in terms of accuracy as well whether it
produces the well-known McGurk illusion; and iii) provide
possible explanations of the neural processing differences
that could lead to a slower maturation of MSI in participants
with ASD, followed by a full recovery during adolescence.
      </p>
      <p>In particular, we describe the training mechanisms
implemented to simulate the maturation phase and we test a
hypothesis to explain ASD deficits in speech MSI: a
different multisensory experience during the maturation
process, due to a lack of attention in young children
(attentional bias) is responsible for the different maturation
in ASD. All the simulated responses are compared with
behavioral data present in the literature.</p>
    </sec>
    <sec id="sec-3">
      <title>Method</title>
      <p>The model consists of a multisensory region (STS) of N
multisensory units (N = 180), receiving excitatory
projections from two arrays of N auditory and N visual units
(see Fig. 1). Unit response to any input is described with a
first order differential equation, which simulates the
integrative properties of the cellular membrane, and a
steady-state sigmoidal relationship, that simulates the
presence of a lower threshold and an upper saturation for
neural activation. The saturation value is set at 1, i.e., all
outputs are normalized to the maximum. In the following,
the term “activity” is used to denote unit output.</p>
      <p>
        Auditory and visual units are devoted to the processing
of information regarding speech sounds and speech gestures
        <xref ref-type="bibr" rid="ref2 ref6">(i.e. lip and face movements; see e.g., Bernstein &amp;
Liebenthal, 2014)</xref>
        , and are topologically organized
according to a similarity principle. This means that two
similar sounds or lips movements activate proximal neural
groups in these areas. The topological organization in these
cortical regions is realized assuming that each unit is
connected with other elements of the same area via lateral
excitatory and inhibitory connections (intra-area
connections, L in Fig. 1), described by a Mexican hat
disposition, i.e., proximal units excite reciprocally and
inhibit more distal ones. This disposition produces an
“activation bubble” in response to a specific auditory or
visual input: not only the neural element representing that
individual feature is activated, but also the proximal ones
linked via sufficient lateral excitation. This arrangement can
have important consequences for the correct perception of
phonemes, for instance resulting in the illusory perceptual
phenomena like the well-known McGurk effect (see section
Results). In this work, lateral intra-area connections are not
subject to training, since we assumed that this process took
place earlier in life than the acquisition of MSI.
      </p>
      <p>
        Furthermore, units in the auditory and visual regions also
receive an external input (corresponding to a speech sound
and/or a gesture representation of the presented phoneme).
These visual and auditory inputs are described with a
gaussian function. The central point of the Gaussian
function corresponds to a specific speech sound/gesture, and
its amplitude with the stimulus intensity; the standard
deviation accounts for the uncertainty of the stimulus
representation. In this model, for simplicity the two inputs
are described with the same function. To reproduce
experimental variability, the external input had been added
with a noisy component, taken from a uniform distribution.
Moreover, since the outside inputs are mediated by
longrange excitatory connections, their temporal aspects are
described by using a second order kinetics, similar to that
commonly adopted to mimic the glutamatergic synaptic
response
        <xref ref-type="bibr" rid="ref11">(i.e., an impulse produces a response similar to an
alpha function, see also Jansen and Rit, 1995)</xref>
        . These
kinetics are characterized by different time constants (a and
v for the two modalities) simulating the auditory and visual
processing in the cortex.
      </p>
      <p>Finally, we consider a cross-modal input, computed
assuming that units of the two areas could be reciprocally
linked via long-range excitatory connections (Wav, Wva, in
Fig 1), described by a pure latency, and the same
secondorder kinetics employed to mimic the temporal aspects of
the external inputs. We assume that, in the network’s initial
configuration, corresponding to an early period of life, these
connections have negligible strength, bur are subject to a
training phase (see below) during which the network learns
to associate the auditory (speech sounds) and visual (speech
gestures) representations of the same phonemes.</p>
      <p>The third area simulates multisensory units in a cortical
region (STS) known to be involved in the phoneme
comprehension tasks, and MSI. These units are linked via
lateral connections with a Mexican-hat arrangement,
implementing a similarity principle (Ls, in Fig. 1).</p>
      <p>Inputs to the multisensory area were generated by
longrange excitatory connections from unisensory regions (Wsv,
Wsa): we used a delayed onset (pure latency) and a
secondorder kinetics to mimic the temporal aspects of these inputs.
The connections between unisensory and multisensory
regions were realized with a Gaussian function, assuming
stronger and more focused connections coming from the
auditory region (Wsa), and more diffuse but weaker
connections coming from the visual area (Wsv). This
asymmetric connectivity helps explain the experimental
results present in the literature about the better abilities in
speech identification in case of auditory stimulation,
compared with the poor performance in the case of visual
inputs. This different representation is assumed being the
final state of a process of unisensory maps refinement in
STS, which takes place in early stages of life. This
development could be included in future implementations of
the model, as an earlier training phase based on the evidence
that auditory stimuli are more informative than the visual
representations of words. The feedforward connectivity is
also responsible for the presence of early weak integrative
phenomena in the younger ASD group (Fig. 1A, Foxe et
al.). In the present model, neither of these connections are
modified during the learning period due to the relative
stability of representations of unisensory speech features at
the ages considered (~7 years of age upward). Finally, the
output of the STS units is compared with a fixed threshold
to mimic the perceptual ability to correctly identify speech
(detection threshold).</p>
    </sec>
    <sec id="sec-4">
      <title>Training the Network</title>
      <p>We simulated a normal training period by presenting
thousands (up to 25.000 inputs) of unisensory and
multisensory speech representations to the network, to
mimic a normal experience with speech stimuli: specifically
we trained the network with 80% of congruent auditory and
visual stimuli and 20% of auditory stimuli alone. During the
training phase we used suprathreshold stimuli at their
highest level of efficacy, i.e. stimuli able to excite
unisensory units close to the upper saturation, in order to
speed up the modeling process.</p>
      <p>
        These stimuli were generated through a uniform
distribution of probability. Each stimulus lasted 130 ms,
during which, after an initial transient period, the
connections among visual and auditory representations of
the same phonemes were crafted by using Hebbian
algorithms of long-term potentiation (LTP) and long-term
depression (LTD). In particular, we chose a presynaptic
gating rule, which means that the training algorithm only
modifies the connections coming from an active unit, and
their strength is modified based on the activity of the
postsynaptic units. As an example, if a presynaptic auditory
element is active, it reinforces connections targeting a
simultaneously active visual unit (likely representing the
same speech unit), and weakens connections with silent
visual elements (likely those coding for different speech
inputs, see Fig. 2)
        <xref ref-type="bibr" rid="ref10 ref19">(see Gerstner, W., &amp; Kistler, 2002)</xref>
        . In
order to establish this correlation, the activity of the
individual units (both presynaptic and postsynaptic) is
compared with a given threshold, to determine whether the
unit can be considered active or silent. The strengthening
and depression processes are also subject to a saturation
rule: which means that each single connection cannot
overcome a maximum value, nor decrease below zero.
      </p>
      <p>Finally, to simulate the delayed developmental processes
taking place in ASD children, we trained and tested the
network by using lower multisensory experiences, precisely
20% multisensory stimuli plus 80% auditory stimuli.
A first set of simulations was performed to evaluate the
network’s ability to correctly identify speech before the
model had been exposed to training (Fig. 3). Already mature
unisensory maps in auditory and visual regions were
supposed in this model, as described in previous section.</p>
      <p>In this phase, representations of speech in the two
unisensory regions are independently activated by the two
modality-specific external stimuli, and do not interact
through direct long-range excitatory projections between the
unisensory cortical regions, which are still ineffective.
Hence, they independently stimulate the corresponding units
in STS region. As shown in Fig. 3, in this initial condition,
an effective auditory stimulus alone is sufficient to produce
a high percentage of correct speech sound identifications, as
in mature adult-like behaviour. If the auditory stimulus is
coupled with a simultaneous visual representation of the
same phoneme, the network shows some benefit, although
this is relatively low and no greater than 20% MSI gain over
all stimuli and levels of efficacy.</p>
      <p>So, the network in its initial stage is characterized by: i)
mature abilities in speech-recognition tasks in case of
auditory-alone stimulation, but ii) poor multisensory
integration (see Fig. 3). These results are in agreement with
what one would expect prior to significant training, and
indeed are well aligned with what we see in our data in
which younger children show relatively immature ability to
benefit from MSI, whereas auditory speech recognition is
significantly closer to mature performance levels.</p>
    </sec>
    <sec id="sec-5">
      <title>Developmental process and audio-visual speech recognition</title>
      <p>The model in its initial state was repeatedly stimulated
with modality-specific and cross-modal inputs (see section
Training) in order to simulate the experience of a child with
different sensory representations of phonemes. The weights
of the inter-area projections among unisensory elements in
the visual and auditory regions adjusted according to
Hebbian dynamics. We tested MSI in the final “adult-like”
configuration and throughout the developmental process,
using the same testing paradigm used to evaluate the MSI
behavior in the immature phase.</p>
      <p>One possible explanation for reduced MSI in ASD is that
learning is less effective in this group. A possible
explanation tested here is that these individuals experience
fewer multisensory exposures, possibly due to how attention
is allocated (e.g., suppression of unattended signals;
selectively focusing on one sensory modality at a time; not
looking at faces consistently). We therefore tested the
impact of percentage of multisensory versus unisensory
exposures on model performance on the maturation of MSI.</p>
      <p>Fig. 4 reports the weight maturation (left panel) and MSI
abilities (right panel) at different epochs, for a training
phase in which the network was exposed to a sensory
training with just 20% of multisensory stimuli.</p>
      <p>Even with such a poor multisensory experience, the
network can reach “TD-like” behaviour in terms of MSI as
shown in Fig. 4, although this maturation requires 15,000
training epochs. This result suggests that multisensory
integration in the model strongly depends on connections
from the visual to the auditory region.</p>
    </sec>
    <sec id="sec-6">
      <title>Simulation of the McGurk effect</title>
      <p>
        An important consequence of training in our model is that
the audio-visual inference becomes stronger after training,
because of connection-weight reinforcement among
unisensory areas. This change have important consequences
in the development of audio-visual illusions. Since
unisensory areas in our model code for speech, a typical
illusion consists in the well-known McGurk effect
        <xref ref-type="bibr" rid="ref17">(McGurk
&amp; McDonald, 1976)</xref>
        . In this illusion, incongruent auditory
speech is dubbed onto visual speech and the resulting
auditory speech percept corresponds to a fusion of the
auditory and visual speech stimuli, or to the visual speech
stimulus, but not to the veridical auditory speech stimulus.
      </p>
      <p>We performed an additional set of simulations with the
model (both in the mature and immature configurations) to
reproduce a McGurk-type situation. Specifically, we
presented mismatched (at four-position distance)
auditoryvisual speech to the network and analyzed the activities
elicited in all areas. We say that the McGurk effect is
evident when the detected phoneme (computed as the
barycenter of activity in the multisensory region) is different
from that used in the auditory input. The network in the
immature configuration is characterized by limited visual
influence on the speech percept. Therefore, the activity in
the auditory region is almost unaffected by the visual
stimulus. In this case, the auditory modality plays the
dominant role in guiding speech perception. In the 42.5% of
presentations, the model identifies the auditory input
correctly, while the McGurk effect is present less than 30%
of the time. In the remaining 27.2% of cases, no phoneme
reaches the detection threshold.</p>
      <p>After training, the model is much more susceptible to the
AV illusion, with responses affected by the visual
information on almost 72% of the simulation trials.</p>
    </sec>
    <sec id="sec-7">
      <title>Discussion</title>
      <p>
        Different computational models have been developed in
recent years to investigate the general problem of
multisensory integration in the brain
        <xref ref-type="bibr" rid="ref25 ref5 ref6 ref7">(see Cuppini et al.,
2011 and Ursino et al., 2014 as a review)</xref>
        . Some of them, in
agreement with several psychophysical and behavioral data,
are based on a Bayesian approach
        <xref ref-type="bibr" rid="ref1 ref12 ref13">(Anastasio et al., 2000;
Knill and Pouget, 2004; Körding et al., 2007)</xref>
        . Others
assume that integration is an emergent property based on
network dynamics
        <xref ref-type="bibr" rid="ref22 ref27">(Patton and Anastasio, 2003; Ursino et
al., 2009)</xref>
        . Finally some models have been realized to deal
with the problem of multisensory integration in semantic
memory and lexical aspects
        <xref ref-type="bibr" rid="ref23 ref24 ref26">(Rogers et al., 2004; Ursino et
al., 2010, 2015)</xref>
        .
      </p>
      <p>
        Concerning the specific problem of speech recognition,
Ma et al
        <xref ref-type="bibr" rid="ref15">(Ma et al., 2009)</xref>
        implemented a Bayesian model of
optimal cue integration able to explain visual influence on
auditory perception in a noisy environment. They explained
different perceptual behaviors based on words
representation as a collection of phonetic features in a
topographically organized feature space.
      </p>
      <p>Although the previous computational efforts simulated
experimental data quite well, none of them was able to
explain the maturation of MSI in speech perception or the
different developmental trajectory for ASD, or how these
capabilities are instantiated in the circuit.</p>
      <p>
        The present model, in its mature architecture, simulates
many experimental findings present in literature regarding
speech MSI. From this point of view, the fundamental
assumption is that the adult configuration implements a
twostep cross-modal integration: the first at the level of
unisensory areas, mediated by the cross-modal connections
between visual and auditory regions; the second at the level
of the multimodal area, due to the presence of convergent
feedforward connections. With this model, we reproduced
the improvement in correct phoneme recognition in
audiovisual vs auditory conditions at different
signal-tonoise levels
        <xref ref-type="bibr" rid="ref9">(Foxe et al., 2015)</xref>
        ; and, we simulated the main
aspects of the McGurk effect.
      </p>
      <p>
        A second important aspect of our study is the capacity to
mimic and to understand the developmental differences
between TD subjects and ASD children regard the
crossmodal abilities observed with age. In particular, the model
explains results of a recent study by
        <xref ref-type="bibr" rid="ref9">Foxe et al. (2015)</xref>
        ,
using two main assumptions. First, the feedforward
connections from unisensory areas to the multisensory area
are already mature in the early age (here, this corresponds to
the condition of the untrained network) and the auditory
feedforward connections are stronger than the visual ones.
Second, the cross-modal connections between unimodal
areas are created during the development, under the pressure
of a multimodal environment (i.e., auditory + visual stimuli)
and this process is faster in TD subjects than in ASDs. This
assumption agrees with the diffuse idea that ASD subjects
have a decreased long-range connectivity, and that autism is
a functional disconnection syndrome, in which the core of
deficit derives from the poor capacity to functionally
connect remote regions of the brain
        <xref ref-type="bibr" rid="ref18">(Melillo &amp; Leisman,
2009)</xref>
        . Since the reason for this decreased connectivity is
still unclear, the model tested a possible scenario where a
reduced number of cross-modal stimuli (reflecting a reduced
attention of the subject to the external world), is a likely
mechanism responsible for the differences in TD and ASD.
      </p>
      <p>These differences may lead to some testable predictions:
from the results about the training phase, when can expect
that ASD children trained with a high percentage of
crossmodal stimuli, could exhibit a normal or at least a quicker
MSI maturation. A second prediction is that as a
consequence of poor cross-modal connections among
unisensory areas, young individual with ASDs have a less
evident McGurk effect, but at the end of the developmental
phase, this illusion becomes comparable in the two classes.
The first prediction is still to be tested; the second is
supported by some experimental results in the literature, in
particular by comparing data across different studies.
However, it deserves a deeper investigation through a
single, ideally longitudinal study.</p>
      <p>Future developments of this model may include a more
detailed and biologically realistic description of the
unisensory areas, and the inclusion of further regions to
simulate the role in MSI and speech perception played by
subcortical structures, like the thalamus and basal ganglia.</p>
    </sec>
    <sec id="sec-8">
      <title>References</title>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Anastasio</surname>
            ,
            <given-names>T.J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Patton</surname>
            ,
            <given-names>P.E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Belkacem-Boussaid</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          (
          <year>2000</year>
          )
          <article-title>Using Bayes rule to model multisensory enhancement in the superior colliculus</article-title>
          .
          <source>Neural Computation</source>
          ,
          <volume>12</volume>
          ,
          <fpage>1165</fpage>
          -
          <lpage>1187</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Bernstein</surname>
            ,
            <given-names>L.E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Liebenthal</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Neural pathways for visual speech perception</article-title>
          .
          <source>Frontiers in neuroscience, 8.</source>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>de Boer-Schellekens</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Keetels</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Eussen</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Vroomen</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          (
          <year>2013</year>
          ).
          <article-title>No evidence for impaired multisensory integration of low-level audiovisual stimuli in adolescents and young adults with autism spectrum disorders</article-title>
          .
          <source>Neuropsychologia</source>
          ,
          <volume>51</volume>
          (
          <issue>14</issue>
          ),
          <fpage>3004</fpage>
          -
          <lpage>3013</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Bolognini</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Miniussi</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Savazzi</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bricolo</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Maravita</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>TMS modulation of visual and auditory processing in the posterior parietal cortex</article-title>
          .
          <source>Experimental brain research</source>
          ,
          <volume>195</volume>
          (
          <issue>4</issue>
          ),
          <fpage>509</fpage>
          -
          <lpage>517</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <surname>Brandwein</surname>
            ,
            <given-names>A.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Foxe</surname>
            ,
            <given-names>J.J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Russo</surname>
            ,
            <given-names>N.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Altschuler</surname>
            ,
            <given-names>T.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Gomes</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Molholm</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>The development of audiovisual multisensory integration across childhood and early adolescence: a high-density electrical mapping study</article-title>
          .
          <source>Cerebral Cortex</source>
          ,
          <volume>21</volume>
          (
          <issue>5</issue>
          ),
          <fpage>1042</fpage>
          -
          <lpage>1055</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bolognini</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vallar</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>A neurocomputational analysis of the sound-induced flash illusion</article-title>
          .
          <source>NeuroImage</source>
          ,
          <volume>92</volume>
          ,
          <fpage>248</fpage>
          -
          <lpage>266</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          (
          <year>2011</year>
          ).
          <article-title>Organization, maturation, and plasticity of multisensory integration: insights from computational modeling studies</article-title>
          .
          <source>Frontiers in psychology, 2.</source>
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Fallon</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Trehub</surname>
            ,
            <given-names>S.E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Schneider</surname>
            ,
            <given-names>B.A.</given-names>
          </string-name>
          (
          <year>2000</year>
          ).
          <article-title>Children's perception of speech in multitalker babble</article-title>
          .
          <source>The Journal of the Acoustical Society of America</source>
          ,
          <volume>108</volume>
          (
          <issue>6</issue>
          ),
          <fpage>3023</fpage>
          -
          <lpage>3029</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Foxe</surname>
            ,
            <given-names>J.J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Molholm</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Del Bene</surname>
            ,
            <given-names>V.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Frey</surname>
            ,
            <given-names>H. P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Russo</surname>
            ,
            <given-names>N.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Blanco</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Saint-Amour</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Ross</surname>
            ,
            <given-names>L.A.</given-names>
          </string-name>
          (
          <year>2015</year>
          ).
          <article-title>Severe multisensory speech integration deficits in highfunctioning school-aged children with autism spectrum disorder (ASD) and their resolution during early adolescence</article-title>
          .
          <source>Cerebral Cortex</source>
          ,
          <volume>25</volume>
          (
          <issue>2</issue>
          ),
          <fpage>298</fpage>
          -
          <lpage>312</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Gerstner</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Kistler</surname>
            ,
            <given-names>W.M.</given-names>
          </string-name>
          (
          <year>2002</year>
          ).
          <article-title>Mathematical formulations of Hebbian learning</article-title>
          .
          <source>Biological cybernetics</source>
          ,
          <volume>87</volume>
          (
          <issue>5-6</issue>
          ),
          <fpage>404</fpage>
          -
          <lpage>415</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Jansen</surname>
            ,
            <given-names>B.H.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Rit</surname>
            ,
            <given-names>V.G.</given-names>
          </string-name>
          (
          <year>1995</year>
          ).
          <article-title>Electroencephalogram and visual evoked potential generation in a mathematical model of coupled cortical columns</article-title>
          .
          <source>Biological cybernetics</source>
          ,
          <volume>73</volume>
          (
          <issue>4</issue>
          ),
          <fpage>357</fpage>
          -
          <lpage>366</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>Knill</surname>
            ,
            <given-names>D.C.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Pouget</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2004</year>
          ).
          <article-title>The Bayesian brain: the role of uncertainty in neural coding and computation</article-title>
          .
          <source>Trends in Neurosciences</source>
          ,
          <volume>27</volume>
          (
          <issue>12</issue>
          ),
          <fpage>712</fpage>
          -
          <lpage>719</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>Körding</surname>
            ,
            <given-names>K.P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Beierholm</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ma</surname>
          </string-name>
          , W.J.,
          <string-name>
            <surname>Quartz</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tenenbaum</surname>
            ,
            <given-names>J.B.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Shams</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          (
          <year>2007</year>
          ).
          <article-title>Causal inference in multisensory perception</article-title>
          .
          <source>PLoS ONE</source>
          <volume>2</volume>
          (
          <issue>9</issue>
          ),
          <year>e943</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Kraus</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Koch</surname>
            ,
            <given-names>D.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>McGee</surname>
            ,
            <given-names>T.J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nicol</surname>
            ,
            <given-names>T.G.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Cunningham</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          (
          <year>1999</year>
          ).
          <article-title>Speech-Sound Discrimination in School-Age Children. Psychophysical and Neurophysiologic Measures</article-title>
          .
          <source>Journal of Speech</source>
          , Language, and Hearing Research,
          <volume>42</volume>
          (
          <issue>5</issue>
          ),
          <fpage>1042</fpage>
          -
          <lpage>1060</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Ma</surname>
          </string-name>
          , W.J.,
          <string-name>
            <surname>Zhou</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ross</surname>
            ,
            <given-names>L.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Foxe</surname>
            ,
            <given-names>J.J.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Parra</surname>
            ,
            <given-names>L.C.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>Lip-reading aids word recognition most in moderate noise: a Bayesian explanation using highdimensional feature space</article-title>
          .
          <source>PLoS One</source>
          ,
          <volume>4</volume>
          (
          <issue>3</issue>
          ),
          <year>e4638</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>A neural network model of ventriloquism effect and aftereffect</article-title>
          .
          <source>PloS one</source>
          ,
          <volume>7</volume>
          (
          <issue>8</issue>
          ),
          <year>e42503</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>McGurk</surname>
            <given-names>H</given-names>
          </string-name>
          ,
          <string-name>
            <surname>MacDonald</surname>
            <given-names>J</given-names>
          </string-name>
          . (
          <year>1976</year>
          ).
          <article-title>Hearing lips and seeing voices</article-title>
          .
          <source>Science</source>
          ,
          <volume>264</volume>
          ,
          <fpage>746</fpage>
          -
          <lpage>8</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <surname>Melillo</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Leisman</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>Autistic spectrum disorders as functional disconnection syndrome</article-title>
          .
          <source>Reviews in the Neurosciences</source>
          ,
          <volume>20</volume>
          (
          <issue>2</issue>
          ),
          <fpage>111</fpage>
          -
          <lpage>132</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <surname>Molholm</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ritter</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Murray</surname>
            ,
            <given-names>M.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Javitt</surname>
            ,
            <given-names>D.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schroeder</surname>
            ,
            <given-names>C. E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Foxe</surname>
            ,
            <given-names>J. J.</given-names>
          </string-name>
          (
          <year>2002</year>
          ).
          <article-title>Multisensory auditory-visual interactions during early sensory processing in humans: a high-density electrical mapping study</article-title>
          .
          <source>Cognitive brain research</source>
          ,
          <volume>14</volume>
          (
          <issue>1</issue>
          ),
          <fpage>115</fpage>
          -
          <lpage>128</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <surname>Molholm</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mercier</surname>
            ,
            <given-names>M.R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Liebenthal</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Schwartz</surname>
            ,
            <given-names>T.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ritter</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Foxe</surname>
            ,
            <given-names>J.J.</given-names>
          </string-name>
          , &amp; De Sanctis,
          <string-name>
            <surname>P.</surname>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Mapping phonemic processing zones along human perisylvian cortex: an electro-corticographic investigation</article-title>
          .
          <source>Brain Structure and Function</source>
          ,
          <volume>219</volume>
          ,
          <fpage>1369</fpage>
          -
          <lpage>1383</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          <string-name>
            <surname>Nath</surname>
            ,
            <given-names>A.R.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Beauchamp</surname>
            ,
            <given-names>M.S.</given-names>
          </string-name>
          (
          <year>2012</year>
          ).
          <article-title>A neural basis for interindividual differences in the McGurk effect, a multisensory speech illusion</article-title>
          .
          <source>Neuroimage</source>
          ,
          <volume>59</volume>
          ,
          <fpage>781</fpage>
          -
          <lpage>787</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <surname>Patton</surname>
            ,
            <given-names>P.E.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Anastasio</surname>
            ,
            <given-names>T.J.</given-names>
          </string-name>
          (
          <year>2003</year>
          ).
          <article-title>Modeling crossmodal enhancement and modality-specific suppression in multisensory neurons</article-title>
          .
          <source>Neural computation</source>
          ,
          <volume>15</volume>
          ,
          <fpage>783</fpage>
          -
          <lpage>810</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <surname>Rogers</surname>
            ,
            <given-names>T.T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lambon</surname>
            <given-names>Ralph</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>M.A.</given-names>
            ,
            <surname>Garrard</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            ,
            <surname>Bozeat</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            ,
            <surname>McClelland</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.L.</given-names>
            ,
            <surname>Hodges</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.R.</given-names>
            , and
            <surname>Patterson</surname>
          </string-name>
          ,
          <string-name>
            <surname>K.</surname>
          </string-name>
          (
          <year>2004</year>
          ).
          <article-title>The structure and deterioration of semantic memory: A neuropsychological and computational investigation</article-title>
          .
          <source>Psychological Review</source>
          ,
          <volume>111</volume>
          ,
          <fpage>205</fpage>
          -
          <lpage>235</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>A computational model of the lexical semantic system based on a grounded cognition approach</article-title>
          . Frontiers in Psychology,
          <volume>1</volume>
          ,
          <fpage>221</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2014</year>
          ).
          <article-title>Neurocomputational approaches to modelling multisensory integration in the brain: A review</article-title>
          .
          <source>Neural Networks</source>
          ,
          <volume>60</volume>
          ,
          <fpage>141</fpage>
          -
          <lpage>165</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , &amp;
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          (
          <year>2015</year>
          ).
          <article-title>A neural network for learning the meaning of objects and words from a featural representation</article-title>
          .
          <source>Neural Networks</source>
          ,
          <volume>63</volume>
          ,
          <fpage>234</fpage>
          -
          <lpage>253</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          <string-name>
            <surname>Ursino</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cuppini</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Magosso</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Serino</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          , &amp; Di Pellegrino,
          <string-name>
            <surname>G.</surname>
          </string-name>
          (
          <year>2009</year>
          ).
          <article-title>Multisensory integration in the superior colliculus: a neural network model</article-title>
          .
          <source>Journal of computational neuroscience</source>
          ,
          <volume>26</volume>
          (
          <issue>1</issue>
          ),
          <fpage>55</fpage>
          -
          <lpage>73</lpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>