<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>How granularity of orthography-phonology mappings affect reading development: Evidence from a computational model of English word reading and spelling</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Alfred Lim and Beth A. O'Brien</string-name>
          <email>alfred.lim@nie.edu.sg</email>
          <email>alfred.lim@nie.edu.sg beth.obrien@nie.edu.sg</email>
          <email>beth.obrien@nie.edu.sg</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Luca Onnis</string-name>
          <email>luca.onnis@unige.it</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Nanyang Technological University</institution>
          ,
          <country country="SG">Singapore</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>University of Genoa</institution>
          ,
          <country country="IT">Italy</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>It is widely held that children implicitly learn the structure of their writing system through statistical learning of spelling-tosound mappings. Yet an unresolved question is how to sequence reading experience so that children can 'pick up' the structure optimally. We tackle this question here using a computational model of encoding and decoding. The order of presentation of words was manipulated so that they exhibited two distinct progressions of granularity of spelling-to-sound mappings. We found that under a training regime that introduced written words progressively from small-to-large granularity, the network exhibited an early advantage in reading acquisition as compared to a regime introducing written words from large-to-small granularity. Our results thus provide support for the grain size theory (Ziegler and Goswami, 2005) and demonstrate that the order of learning can influence learning trajectories of literacy skills.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1 Introduction</title>
      <p>
        Reading science provides evidence of the
developmental path to acquiring reading for alphabetic
languages
        <xref ref-type="bibr" rid="ref28 ref7">(Ehri, 2005; Rayner et al., 2001)</xref>
        . From
parsing the speech stream into words in infancy
        <xref ref-type="bibr" rid="ref31 ref4">(Christiansen et al., 2006; Saffran et al., 1997)</xref>
        ,
to familiarizing with print in the preschool years
        <xref ref-type="bibr" rid="ref33">(Thompson, 2009)</xref>
        — these activities lead to the
accrual of key knowledge for learning to read.
Knowledge about the language’s phonotactic and
      </p>
      <p>Copyright ©2020 for this paper by its authors. Use
permitted under Creative Commons License Attribution 4.0
International (CC BY 4.0).
graphotactic properties and symbolic
representations with abstract letter units is necessary for the
forthcoming insight that print represents spoken
language (the alphabetic principle). Subsequent
to this insight, children are ready to take on the
process of learning the precise mapping of
printto-speech.</p>
      <p>
        At its basis, learning to read involves
learning to decode a script into oral language
representations. The question arises as to the optimal
input for learning this orthography-to-phonology
mapping in an alphabetic system, especially for
languages that have deep orthographies, such as
English. Shallow orthographies (e.g., Finnish,
Spanish) have a more precise match between
letters and sounds; whereas deep orthographies
match phonemes to graphemes (one or more
letters) in an inconsistent way — with multiple
spellings per phoneme, or multiple pronunciations
per grapheme — and, thus, have a greater number
of GPCs (grapheme-phoneme correspondences).
Therefore, reading acquisition is found to occur
at a comparatively slower rate for readers in deep
as compared to shallow orthographies
        <xref ref-type="bibr" rid="ref10 ref32 ref8 ref9">(Ellis et
al., 2004; Georgiou et al., 2008; Florit and Cain,
2011)</xref>
        .
      </p>
      <p>
        The deep orthographic complexity of English
also partly results from variation in the functional
units of the writing system — graphemes
—which may consist of a single letter (e.g., a), or
multiple letters (e.g., ay, aye). While skilled adult
readers have unitized these subword patterns
        <xref ref-type="bibr" rid="ref29">(Rey
et al., 2000)</xref>
        , beginning readers need to acquire
these patterns of graphemes and their mappings to
phonemes. Here we consider this mapping
problem along two dimensions: (1) the granularity of
the units of analysis to be picked up at any given
time; and (2) the ordering of learning such units
and types.
      </p>
      <p>
        A fruitful approach to examining the GPC
learning process is through computational
modelling
        <xref ref-type="bibr" rid="ref19 ref23 ref24 ref26">(Monaghan and Ellis, 2010; Perry et al.,
2019; Pritchard et al., 2016)</xref>
        . Specifically,
connectionist models are sensitive to the timing and
ordering of learning events, in that they learn
incrementally. This feature is particularly apt for
modeling reading development, as it affords simulating
the incremental nature of a child learning to read
new words daily, as schooling progresses. Order
effects as well as frequency trajectory effects have
been documented in previous connectionist
models
        <xref ref-type="bibr" rid="ref18">(Mermillod et al., 2012)</xref>
        , and here we are
interested in comparing learning trajectories for
particular training orderings for reading development.
      </p>
      <p>
        To this end, we present connectionist networks
with small batches of words, which we test
regularly for accuracy until a given criterion across
the batch is achieved — in essence, an adaptive
training regime. Using this approach, we can
address long-standing issues in the area of
reading education with a more systematic approach to
understanding how print-to-speech mappings are
learned
        <xref ref-type="bibr" rid="ref30">(Rueckl, 2016)</xref>
        .
      </p>
      <p>Below we briefly review why print-to-speech
decoding can be a hard problem, both for learners
and for researchers trying to understand its
mechanisms. Then, we discuss dimensions of granularity
derived from the literature, and offer a first set of
connectionist simulations of the order of reading
acquisition of American English.</p>
      <sec id="sec-1-1">
        <title>1.1 Is there an optimal reading experience?</title>
        <p>
          The psycholinguistic grain size theory
          <xref ref-type="bibr" rid="ref41">(Ziegler
and Goswami, 2005)</xref>
          has generated much research
on reading acquisition, including across
different alphabetic languages. It espouses that
granularity for oral and written language development
proceed in different directions — from larger to
smaller, vs. from smaller to larger units. Thus, the
mismatch in unit or “grain” sizes available over
development introduces a disparity in learning the
mapping between orthography and phonology.
        </p>
        <p>
          This learning challenge has led to investigations
of behavioral interventions for teaching reading
at either whole word or subword levels
          <xref ref-type="bibr" rid="ref16 ref21">(National
Reading Panel, 2001; McArthur et al., 2015)</xref>
          ,
showing an advantage for subword approaches
emphasizing letter, grapheme or larger
(subsyllable onset-rime) units
          <xref ref-type="bibr" rid="ref22 ref28 ref34 ref5 ref6">(Rayner et al., 2001; Ehri et
al., 2001; Torgerson et al., 2006; Olson and Wise,
1992; Ecalle et al., 2009)</xref>
          . At the same time, the
optimal subword grain size has been debated.
Developmentally, Treiman et al. (2006) reported that
children appear to initially attend to small units
(graphemes), before gradually showing an
influence of surrounding graphemes when confronted
with inconsistencies in pronunciation.
        </p>
        <p>Thus, in the current study we focus on
single grapheme to phonemes, or single phoneme to
grapheme mappings in our inquiry of
granularity and learning to read. In this way, we make
no assumptions about a beginning reader’s
knowledge of subword units or syllable structure, instead
assuming all letters are created equal (whether
vowels or consonants) and that the reading
system must initially acquire knowledge of print
patterns for GPC on its own, through experience
with the print input. Granularity was,
therefore, operationalised for each word as the
difference between the number of letters (Nletter) and
phonemes (Nphon; i.e., Nletter Nphon). For
example, the granularity of the word mince (Nletter
= 5, Nphon = 4) is 1, and the granularity of thought
(Nletter = 7, Nphon = 3) is 4. A granularity of
0, hence, indicates that the word comprises of no
multi-letter grapheme (e.g., held, storm).</p>
        <p>
          The aim of this study was to systematically
examine granularity related to the learning of GPC
and word decoding. Theoretical accounts of the
best representational units for learning to read
have not been explicitly tested in the modelling
literature to our knowledge. This, in turn, may
inform instructional practices as to the best
approaches for optimizing the learning curve, and
results can be interpreted in terms of optimal child
developmental trajectories and reading curricula
          <xref ref-type="bibr" rid="ref17">(McKeown et al., 2017)</xref>
          .
2
2.1
        </p>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>Method</title>
      <sec id="sec-2-1">
        <title>Model Architecture</title>
        <p>
          The model had four types of layers: orthographic,
phonological, hidden, and clean-up (see Figure 1).
The orthographic and phonological layers were
each connected to a clean-up layer that mediated
connections within the respective units, creating
an attractor network that settles into a stable
pattern over time
          <xref ref-type="bibr" rid="ref11">(Harm and Seidenberg, 1999)</xref>
          .
        </p>
        <p>The orthographic layer was composed of 260
units, corresponding to 10 positions 26
possible letters. Words were coded as vowel-centred,
such that the fourth slot was filled with the
leftmost vowel of a word (e.g., mince ! m i n c e
).</p>
        <p>
          A word’s phonology was represented with
nodes coding features of phonemes (8 positions
28 possible phonological features = 224 units).
Pronunciation of each word was positioned with
the vowel at the fourth slot (e.g., mince ! /
m I n s /). Each phoneme was encoded by a
binary vector of 28 phonological features taken from
PHOIBLE
          <xref ref-type="bibr" rid="ref20 ref24">(Moran and McCloy, 2019)</xref>
          , an online
repository of cross-lingual phonological data. A
list of phonemes and their respective phonological
features used in the present work can be found in
the Open Science Framework (OSF) repository for
this project (https://osf.io/hj96x/).
        </p>
        <p>While traditionally the problem of learning to
read is conceptualized in terms of decoding
unidirectionally from orthography to phonology,
research suggests that children engage in spelling
words simultaneously as they learn to decode.
In addition, feedback sound-to-spelling relations
are also informative in establishing mappings for
reading. Thus, we implemented a new model with
a bidirectional network architecture that connects
orthographic-to-phonological and
phonologicalto-orthographic layers via the hidden units.
2.2</p>
      </sec>
      <sec id="sec-2-2">
        <title>Training Procedure</title>
        <p>
          The model was trained with a learning rate of 0.05
using a back-propagation through time (BPTT)
algorithm with input integration and a time constant
of 0.5
          <xref ref-type="bibr" rid="ref11 ref25">(Harm and Seidenberg, 1999; Plaut et al.,
1996)</xref>
          . Each word item was clamped and
presented for six time ticks, and then in an additional
six time ticks, the model was required to
reproduce the target pattern of the word by the final 12th
tick. The weight connections are updated based on
cross-entropy error computed between the target
and the actual activation of the output units.
        </p>
        <p>Training proceeded in two distinct stages
reflecting naturalistic child language development:
(1) a pre-literacy training stage, in which the
model was trained to learn the
phonologyto-phonology mappings with an accuracy of
99%; and (2) a literacy training stage, in
which the model was trained on both
decoding (orthography-to-phonology) and encoding
(phonology-to-orthography) tasks in a sequential
manner. The pre-literacy stage of training was
intended to mimic the fact that children develop oral
skills through hearing and speaking long before
learning to read.</p>
        <p>Models were trained with a cumulative
process of learning to encode and decode, whereby
words with different granularity were introduced
to the model in either an ascending or
descending sequence. These two models were referred to
as small–to–large (SL) and large–to–small (LS)
from here onwards.</p>
        <p>
          Words were first sorted with regard to their
granularity, followed by a second-level sorting
criterion to arrange words with the same granularity
in order of decreasing frequency. The first batch
of words in each training regime, therefore,
comprises of high frequency words that are of either
the smallest (e.g., fix, lynx) or largest (e.g., bought,
should) granularity in the corpus. During training,
words were sampled according to their frequency
from the Word Frequency Guide (WFG) corpus
          <xref ref-type="bibr" rid="ref40">(Zeno et al., 1995)</xref>
          , and the resulting probability
values were normalized over all words in the
training set. Correspondingly, low frequency words
had a lower probability of being presented to the
model during training as compared to high
frequency words [e.g., P (yules) = 0:05 vs. P (of) =
0:97].
        </p>
      </sec>
      <sec id="sec-2-3">
        <title>2.2.1 Adaptive training</title>
        <p>
          Teachers introduce written words progressively
to their pupils, and regularly assess progress
before introducing new words. Likewise, our model
training introduced batches of 45 new words at a
time. Importantly, a new batch of words was
introduced only after model performance exceeded a
criterion threshold of 70% combined accuracy for
the decoding and encoding tasks on trained words
— which included only words that the model had
been trained on cumulatively up to the last training
epoch. This tested the network success at
reproducing the training set to which it had been
progressively exposed, and allowed us to compare the
rates of learning under different training regimes.
Two complementary tests are carried out every
100 training epochs: (1) a total vocabulary test
which uses words from the entire corpus,
regardless of whether they have been presented to the
model in previous training phases; and (2) an
untrained pseudo-words test which uses a fixed set
of pronounceable and spellable monosyllabic
nonwords. This pseudo-word set is derived from
previous empirical studies on developmental reading
skills
          <xref ref-type="bibr" rid="ref35">(Torgesen et al., 1999)</xref>
          . Thus, with these
tests we assess the network’s (1) transfer and (2)
decoding abilities.
        </p>
        <p>Because no learning occurs during testing, the
same set of test words and non-words can be used
routinely as novel testing items after 100 training
epochs. This represents a considerable advantage
with respect to behavioural longitudinal
experiments, where successive test sessions can suffer
from previous exposure effects.</p>
        <p>Each test was administered twice, once in a
decoding task and again in an encoding task. The
decoding task activated the orthographic pattern for
a given test word on the orthographic layer, say,
eye, and measured the accuracy of the network to
reproduced the corresponding target phonological
word (/aI/) on the phonological layer. Conversely,
the encoding task activated the phonological
pattern for a given word on the phonology layer, say,
/aI/, and measured the accuracy of the network to
reproduced the corresponding target orthographic
word (eye) on the orthographic layer.</p>
        <p>Similar to the training procedure, each test word
item was clamped and presented for six time
samples, and then in an additional six time
samples, the model was required to produce the
target phonological/orthographic pattern of the word.
An output was scored as correct when the target
nodes were active with a value &gt;= 0.75, and
concurrently the other nodes were inactive (&lt;= 0.25).
Intermediate values were considered incorrect.
2.4</p>
      </sec>
      <sec id="sec-2-4">
        <title>Corpus</title>
        <p>
          All stimuli were monosyllabic American English
words. The CHILDES database
          <xref ref-type="bibr" rid="ref15">(MacWhinney,
2000)</xref>
          , WFG corpus
          <xref ref-type="bibr" rid="ref40">(Zeno et al., 1995)</xref>
          , and
the Phonemic Decoding Efficiency sub-test of the
TOWRE
          <xref ref-type="bibr" rid="ref35">(Torgesen et al., 1999)</xref>
          were used for
preliteracy training (N = 5032), literacy training (N
= 4394), and pseudo-words testing (N = 163),
respectively. The full list of words and their
respective granularity and batch number can be found on
OSF (https://osf.io/hj96x/).
        </p>
        <p>To check whether frequency covaries with grain
size, and may therefore confound the order
effect, we conducted Spearman’s correlations across
the training regime between batch number and
mean log frequency per batch. This was done
for each training order: small–to–large (SL) and
large–to–small (LS). Importantly, while batch
number was significantly correlated with
frequency for both training orders [SL: rs(96) = -0.43,
p &lt; .001; LS: rs(96) = -0.30, p = .003], the relation
was in the same, negative direction in both cases
—- ensuring that frequency was not systematically
tied to grain size. Rather, the result was from the
second-level sorting by frequency in descending
order.</p>
        <p>
          To identify the possible relationship between
the granularity and consistency of the mapping
for the units to be learned, we calculated the
decoding and encoding consistency measures to
reflect how often the orthographic/phonological unit
was spelled/pronounced in the same way as it was
across all words
          <xref ref-type="bibr" rid="ref3">(Berndt et al., 1987)</xref>
          . The
procedure required the conditional probabilities of
GPCs and PGCs to be computed as they occur in
the corpus [e.g., the probability of the grapheme
ew being pronounced as /o/ is, P (/o/jew) =
0:057].
        </p>
        <p>We then derived a composite consistency score
to account for the two measures (decoding and
encoding), with a higher score representing higher
overall bi-directional word consistency.
Consistency was found to correlate negatively with
granularity increases [SL: rs(96) = -0.69, p &lt; .001; LS:
rs(96) = 0.71, p &lt; .001], indicating that words with
smaller granularity were more consistent.
3</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Results</title>
      <p>At the time of writing, each model had been
trained on 67 out of 98 batches of words (or 3015
out of 4394 unique words). While incomplete,
our preliminary observations suggest a clear
difference in rate of learning across the two training
regimes.</p>
      <p>Results are summarised in Figure 2, and show
that under a training regime that introduces
written words in batches progressively from
small-tolarge granularity, the network exhibited an early
advantage in reading acquisition as compared to
a regime introducing written words from
large-tosmall granularity.</p>
      <p>The two types of repeated tests served to
evaluate the accuracy of phonological output for: (a)
total vocabulary (including trained and untrained
words) and (b) pseudo-words. Both tests
measured the ability of the networks to generalize to
unseen but orthographically legal strings (see
Figure 2). Specifically, the SL and LS models took
232800 and 346400 epochs, respectively, to reach
the criterion threshold of 70% accuracy for all
67 batches of words that were introduced
cumulatively over time. Apart from reaching the
criterion threshold earlier, the SL model also
performed better than the LS model in pseudo-words
test (47.85% vs. 33.13%) at the end of preliminary
training.
4</p>
    </sec>
    <sec id="sec-4">
      <title>Discussion</title>
      <p>
        As the process of learning to read requires picking
up and internalizing representational units of print
associated with sound, the ordering of training
input to the reading system becomes paramount.
How best to order input and maximize learning
efficiency has been debated in the literacy education
field. This study capitalizes on a computational
modelling approach to this issue, using a highly
controlled context without the ethical concerns of
human learning studies. Directly contrasting the
effects of two literacy training regimes differing
in granularity order, the simulation results support
better learning with smaller, less complex
orthographic units, as predicted from corpus-based
research
        <xref ref-type="bibr" rid="ref38">(Vousden, 2008)</xref>
        . At training stages
comprising of 3015 words, we found that the model
initially trained with words of smaller granularity
performed and generalized to pseudo-words better
than the model trained with larger granularity. The
LS model did require significantly more training
epochs to reach the same performance as the SL
model.
      </p>
      <p>
        Essentially, when children learn to read, they
must navigate the structure of their language and
its writing system. Granularity and consistency
are important aspects of this structure, and both
impact reading performance. Adult readers are
slower to identify letters within a multi-letter
grapheme
        <xref ref-type="bibr" rid="ref29 ref32 ref9">(Smith and Monaghan, 2011; Rey et al.,
2000)</xref>
        , suggesting that graphemes are functional
reading units. Furthermore, Rastle and Coltheart
(1998) found that naming latencies were slower
for pseudo-words with, as compared to without,
multi-letter graphemes. Adult word naming and
lexical decision are also faster for consistent words
        <xref ref-type="bibr" rid="ref12 ref13 ref2">(Andrews, 1982; Jared, 1997; Jared, 2002)</xref>
        , and
consistent words are more accurately read and
spelled by children
        <xref ref-type="bibr" rid="ref1 ref14 ref39">(Alegria and Mousty, 1996;
Le´te´ et al., 2008; Weekes et al., 2006)</xref>
        .
      </p>
      <p>
        Granularity and consistency have been regarded
to be associated
        <xref ref-type="bibr" rid="ref36">(Treiman et al., 1995)</xref>
        , and our
corpus analysis revealed this as well —
monosyllabic English words of smaller granularity tend
to be more consistent than words with larger
granularity. This relationship indicates that
granularity and consistency may not be entirely
disentangled, at least for English. With this in mind, the
SL model was first exposed to words of smaller
granularity that were also more consistent in their
GPC and PGC (phoneme-grapheme
correspondence) mappings. Thus consistency and
granularity may be two sides of the same coin, and
when manipulated they could lead to faster or
slower rates of convergence. Importantly, the
current model included bidirectional links between
orthographic and phonological units, simulating
the real-world scenario that children acquire
decoding and encoding skills simultaneously.
      </p>
      <p>
        These findings have implications for
educational planning for early literacy. In particular,
our pilot simulation provides preliminary evidence
on the potential utility of manipulating the order
of training in terms of word granularity to unveil
facilitative effects on literacy acquisition.
Reading instruction can consider the early acquisition
of words with smaller granularity, or more
consistency. However, we note that the present findings
are based on the analysis of monosyllabic words
only and should not be generalized to
multisyllabic words directly. Future work can consider
using models that are capable of reading
multisyllabic words
        <xref ref-type="bibr" rid="ref23">(Perry et al., 2010)</xref>
        , or explore the
link between granularity and consistency across
languages that are either less or more
orthographically transparent.
      </p>
    </sec>
    <sec id="sec-5">
      <title>Acknowledgments</title>
      <p>Support comes from the Education Research
Funding Programme of the National Institute of
Education (NIE), Nanyang Technological
University, Singapore, grant #OER0417OBA.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <given-names>Jesus</given-names>
            <surname>Alegria</surname>
          </string-name>
          and
          <string-name>
            <given-names>Philippe</given-names>
            <surname>Mousty</surname>
          </string-name>
          .
          <year>1996</year>
          .
          <article-title>The development of spelling procedures in french-speaking, normal and reading-disabled children: Effects of frequency and lexicality</article-title>
          .
          <source>Journal of experimental child psychology</source>
          ,
          <volume>63</volume>
          (
          <issue>2</issue>
          ):
          <fpage>312</fpage>
          -
          <lpage>338</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <given-names>Sally</given-names>
            <surname>Andrews</surname>
          </string-name>
          .
          <year>1982</year>
          .
          <article-title>Phonological recoding: Is the regularity effect consistent?</article-title>
          <source>Memory &amp; Cognition</source>
          ,
          <volume>10</volume>
          (
          <issue>6</issue>
          ):
          <fpage>565</fpage>
          -
          <lpage>575</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Rita S. Berndt</surname>
          </string-name>
          , James A.
          <string-name>
            <surname>Reggia</surname>
          </string-name>
          , and
          <string-name>
            <surname>Charlotte</surname>
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Mitchum</surname>
          </string-name>
          .
          <year>1987</year>
          .
          <article-title>Empirically derived probabilities for grapheme-to-phoneme correspondences in english</article-title>
          .
          <source>Behavior Research Methods</source>
          , Instruments, &amp;
          <string-name>
            <surname>Computers</surname>
          </string-name>
          ,
          <volume>19</volume>
          (
          <issue>1</issue>
          ):
          <fpage>1</fpage>
          -
          <lpage>9</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Morten H. Christiansen</surname>
          </string-name>
          , Stephen A.
          <string-name>
            <surname>Hockema</surname>
            , and
            <given-names>Luca</given-names>
          </string-name>
          <string-name>
            <surname>Onnis</surname>
          </string-name>
          .
          <year>2006</year>
          .
          <article-title>Using phoneme distributions to discover words and lexical categories in unsegmented speech</article-title>
          .
          <source>In Proceedings of the Annual Meeting of the Cognitive Science Society</source>
          , volume
          <volume>28</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <given-names>Jean</given-names>
            <surname>Ecalle</surname>
          </string-name>
          , Annie Magnan, and
          <string-name>
            <given-names>Caroline</given-names>
            <surname>Calmus</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>Lasting effects on literacy skills with a computer-assisted learning using syllabic units in low-progress readers</article-title>
          .
          <source>Computers &amp; Education</source>
          ,
          <volume>52</volume>
          (
          <issue>3</issue>
          ):
          <fpage>554</fpage>
          -
          <lpage>561</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <given-names>Linnea C.</given-names>
            <surname>Ehri</surname>
          </string-name>
          ,
          <string-name>
            <surname>Simone R. Nunes</surname>
          </string-name>
          , Steven A.
          <string-name>
            <surname>Stahl</surname>
          </string-name>
          , and
          <string-name>
            <surname>Dale</surname>
            <given-names>M.</given-names>
          </string-name>
          <string-name>
            <surname>Willows</surname>
          </string-name>
          .
          <year>2001</year>
          .
          <article-title>Systematic phonics instruction helps students learn to read: Evidence from the national reading panel's meta-analysis</article-title>
          .
          <source>Review of Educational Research</source>
          ,
          <volume>71</volume>
          (
          <issue>3</issue>
          ):
          <fpage>393</fpage>
          -
          <lpage>447</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <given-names>Linnea C.</given-names>
            <surname>Ehri</surname>
          </string-name>
          .
          <year>2005</year>
          .
          <article-title>Learning to read words: Theory, findings, and issues</article-title>
          .
          <source>Scientific Studies of reading</source>
          ,
          <volume>9</volume>
          (
          <issue>2</issue>
          ):
          <fpage>167</fpage>
          -
          <lpage>188</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <given-names>Nick C.</given-names>
            <surname>Ellis</surname>
          </string-name>
          , Miwa Natsume, Katerina Stavropoulou, Lorenc Hoxhallari,
          <string-name>
            <surname>Victor H. P. Van Daal</surname>
          </string-name>
          ,
          <string-name>
            <surname>Nicoletta Polyzoe</surname>
          </string-name>
          ,
          <string-name>
            <surname>Maria-Louisa Tsipa</surname>
            , and
            <given-names>Michalis</given-names>
          </string-name>
          <string-name>
            <surname>Petalas</surname>
          </string-name>
          .
          <year>2004</year>
          .
          <article-title>The effects of orthographic depth on learning to read alphabetic, syllabic, and logographic scripts</article-title>
          . Reading research quarterly,
          <volume>39</volume>
          (
          <issue>4</issue>
          ):
          <fpage>438</fpage>
          -
          <lpage>468</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <given-names>Elena</given-names>
            <surname>Florit</surname>
          </string-name>
          and
          <string-name>
            <given-names>Kate</given-names>
            <surname>Cain</surname>
          </string-name>
          .
          <year>2011</year>
          .
          <article-title>The simple view of reading: Is it valid for different types of alphabetic orthographies? Educational Psychology Review</article-title>
          ,
          <volume>23</volume>
          (
          <issue>4</issue>
          ):
          <fpage>553</fpage>
          -
          <lpage>576</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>George K. Georgiou</surname>
          </string-name>
          , Rauno Parrila, and
          <string-name>
            <surname>Timothy</surname>
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Papadopoulos</surname>
          </string-name>
          .
          <year>2008</year>
          .
          <article-title>Predictors of word decoding and reading fluency across languages varying in orthographic consistency</article-title>
          .
          <source>Journal of Educational Psychology</source>
          ,
          <volume>100</volume>
          (
          <issue>3</issue>
          ):
          <fpage>566</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <given-names>Michael W.</given-names>
            <surname>Harm</surname>
          </string-name>
          and
          <string-name>
            <given-names>Mark S.</given-names>
            <surname>Seidenberg</surname>
          </string-name>
          .
          <year>1999</year>
          .
          <article-title>Phonology, reading acquisition, and dyslexia: insights from connectionist models</article-title>
          .
          <source>Psychological review</source>
          ,
          <volume>106</volume>
          (
          <issue>3</issue>
          ):
          <fpage>491</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <given-names>Debra</given-names>
            <surname>Jared</surname>
          </string-name>
          .
          <year>1997</year>
          .
          <article-title>Spelling-sound consistency affects the naming of high-frequency words</article-title>
          .
          <source>Journal of Memory and Language</source>
          ,
          <volume>36</volume>
          (
          <issue>4</issue>
          ):
          <fpage>505</fpage>
          -
          <lpage>529</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <given-names>Debra</given-names>
            <surname>Jared</surname>
          </string-name>
          .
          <year>2002</year>
          .
          <article-title>Spelling-sound consistency and regularity effects in word naming</article-title>
          .
          <source>Journal of Memory and Language</source>
          ,
          <volume>46</volume>
          (
          <issue>4</issue>
          ):
          <fpage>723</fpage>
          -
          <lpage>750</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <given-names>Bernard</given-names>
            <surname>Le</surname>
          </string-name>
          ´te´,
          <source>Ronald Peereman, and Michel Fayol</source>
          .
          <year>2008</year>
          .
          <article-title>Consistency and word-frequency effects on spelling among first-to fifth-grade french children: A regression-based study</article-title>
          .
          <source>Journal of Memory and Language</source>
          ,
          <volume>58</volume>
          (
          <issue>4</issue>
          ):
          <fpage>952</fpage>
          -
          <lpage>977</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <given-names>Brian</given-names>
            <surname>MacWhinney</surname>
          </string-name>
          .
          <year>2000</year>
          .
          <article-title>The CHILDES project: The database</article-title>
          , volume
          <volume>2</volume>
          . Psychology Press.
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          <string-name>
            <surname>Genevieve</surname>
            <given-names>McArthur</given-names>
          </string-name>
          ,
          <string-name>
            <given-names>Saskia</given-names>
            <surname>Kohnen</surname>
          </string-name>
          , Kristy Jones, Philippa Eve, Erin Banales, Linda Larsen, and
          <string-name>
            <given-names>Anne</given-names>
            <surname>Castles</surname>
          </string-name>
          .
          <year>2015</year>
          .
          <article-title>Replicability of sight word training and phonics training in poor readers: a randomised controlled trial</article-title>
          .
          <source>PeerJ</source>
          ,
          <volume>3</volume>
          :
          <fpage>e922</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          <string-name>
            <surname>Margaret G. McKeown</surname>
            ,
            <given-names>Paul D.</given-names>
          </string-name>
          <string-name>
            <surname>Deane</surname>
          </string-name>
          , and
          <string-name>
            <surname>Ren</surname>
            <given-names>R.</given-names>
          </string-name>
          <string-name>
            <surname>Lawless</surname>
          </string-name>
          .
          <year>2017</year>
          .
          <article-title>Vocabulary assessment to support instruction: Building rich word-learning experiences</article-title>
          .
          <source>Guilford Publications.</source>
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          <string-name>
            <given-names>Martial</given-names>
            <surname>Mermillod</surname>
          </string-name>
          , Patrick Bonin, Alain Me´ot, Ludovic Ferrand, and
          <string-name>
            <given-names>Michel</given-names>
            <surname>Paindavoine</surname>
          </string-name>
          .
          <year>2012</year>
          .
          <article-title>Computational evidence that frequency trajectory theory does not oppose but emerges from age-ofacquisition theory</article-title>
          .
          <source>Cognitive Science</source>
          ,
          <volume>36</volume>
          (
          <issue>8</issue>
          ):
          <fpage>1499</fpage>
          -
          <lpage>1531</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          <string-name>
            <given-names>Padraic</given-names>
            <surname>Monaghan</surname>
          </string-name>
          and
          <string-name>
            <given-names>Andrew W.</given-names>
            <surname>Ellis</surname>
          </string-name>
          .
          <year>2010</year>
          .
          <article-title>Modeling reading development: Cumulative, incremental learning in a computational model of word naming</article-title>
          .
          <source>Journal of Memory and Language</source>
          ,
          <volume>63</volume>
          (
          <issue>4</issue>
          ):
          <fpage>506</fpage>
          -
          <lpage>525</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          <string-name>
            <given-names>Steven</given-names>
            <surname>Moran</surname>
          </string-name>
          and Daniel McCloy, editors.
          <source>2019. PHOIBLE 2</source>
          .0. Jena:
          <article-title>Max Planck Institute for the Science of Human History</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          National Reading Panel.
          <year>2001</year>
          .
          <article-title>Report of the National Reading Panel: Teaching Children to Read (Report NIH Pub</article-title>
          . No.
          <volume>00</volume>
          -
          <fpage>4769</fpage>
          ).
          <source>National Institutes of Health.</source>
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          <string-name>
            <given-names>Richard K.</given-names>
            <surname>Olson</surname>
          </string-name>
          and
          <string-name>
            <given-names>Barbara W.</given-names>
            <surname>Wise</surname>
          </string-name>
          .
          <year>1992</year>
          .
          <article-title>Reading on the computer with orthographic and speech feedback</article-title>
          .
          <source>Reading and Writing</source>
          ,
          <volume>4</volume>
          (
          <issue>2</issue>
          ):
          <fpage>107</fpage>
          -
          <lpage>144</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          <string-name>
            <given-names>Conrad</given-names>
            <surname>Perry</surname>
          </string-name>
          , Johannes C. Ziegler, and
          <string-name>
            <given-names>Marco</given-names>
            <surname>Zorzi</surname>
          </string-name>
          .
          <year>2010</year>
          .
          <article-title>Beyond single syllables: Large-scale modeling of reading aloud with the connectionist dual process (cdp++) model</article-title>
          . Cognitive psychology,
          <volume>61</volume>
          (
          <issue>2</issue>
          ):
          <fpage>106</fpage>
          -
          <lpage>151</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          <string-name>
            <given-names>Conrad</given-names>
            <surname>Perry</surname>
          </string-name>
          , Marco Zorzi, and
          <string-name>
            <surname>Johannes</surname>
            <given-names>C.</given-names>
          </string-name>
          <string-name>
            <surname>Ziegler</surname>
          </string-name>
          .
          <year>2019</year>
          .
          <article-title>Understanding dyslexia through personalized large-scale computational models</article-title>
          .
          <source>Psychological science</source>
          ,
          <volume>30</volume>
          (
          <issue>3</issue>
          ):
          <fpage>386</fpage>
          -
          <lpage>395</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          <string-name>
            <given-names>David C.</given-names>
            <surname>Plaut</surname>
          </string-name>
          ,
          <string-name>
            <surname>James L. McClelland</surname>
            ,
            <given-names>Mark S.</given-names>
          </string-name>
          <string-name>
            <surname>Seidenberg</surname>
            , and
            <given-names>Karalyn</given-names>
          </string-name>
          <string-name>
            <surname>Patterson</surname>
          </string-name>
          .
          <year>1996</year>
          .
          <article-title>Understanding normal and impaired word reading: computational principles in quasi-regular domains</article-title>
          .
          <source>Psychological review</source>
          ,
          <volume>103</volume>
          (
          <issue>1</issue>
          ):
          <fpage>56</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          <string-name>
            <given-names>Stephen C.</given-names>
            <surname>Pritchard</surname>
          </string-name>
          , Max Coltheart, Eva Marinus, and
          <string-name>
            <given-names>Anne</given-names>
            <surname>Castles</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Modelling the implicit learning of phonological decoding from training on whole-word spellings and pronunciations</article-title>
          .
          <source>Scientific studies of reading</source>
          ,
          <volume>20</volume>
          (
          <issue>1</issue>
          ):
          <fpage>49</fpage>
          -
          <lpage>63</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref27">
        <mixed-citation>
          <string-name>
            <given-names>Kathleen</given-names>
            <surname>Rastle</surname>
          </string-name>
          and
          <string-name>
            <given-names>Max</given-names>
            <surname>Coltheart</surname>
          </string-name>
          .
          <year>1998</year>
          .
          <article-title>Whammy and double whammy: Length effects in nonword naming</article-title>
          .
          <source>Psychonomic Bulletin and Review</source>
          ,
          <volume>5</volume>
          :
          <fpage>277</fpage>
          -
          <lpage>282</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref28">
        <mixed-citation>
          <string-name>
            <given-names>Keith</given-names>
            <surname>Rayner</surname>
          </string-name>
          ,
          <string-name>
            <surname>Barbara R. Foorman</surname>
          </string-name>
          , Charles A.
          <string-name>
            <surname>Perfetti</surname>
          </string-name>
          , David Pesetsky, and Mark S Seidenberg.
          <year>2001</year>
          .
          <article-title>How psychological science informs the teaching of reading. Psychological science in the public interest, 2(2</article-title>
          ):
          <fpage>31</fpage>
          -
          <lpage>74</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref29">
        <mixed-citation>
          <string-name>
            <given-names>Arnaud</given-names>
            <surname>Rey</surname>
          </string-name>
          , Johannes C. Ziegler, and
          <string-name>
            <surname>Arthur</surname>
            <given-names>M.</given-names>
          </string-name>
          <string-name>
            <surname>Jacobs</surname>
          </string-name>
          .
          <year>2000</year>
          .
          <article-title>Graphemes are perceptual reading units</article-title>
          .
          <source>Cognition</source>
          ,
          <volume>75</volume>
          (
          <issue>1</issue>
          ):
          <fpage>B1</fpage>
          -
          <lpage>B12</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref30">
        <mixed-citation>
          <string-name>
            <given-names>Jay G.</given-names>
            <surname>Rueckl</surname>
          </string-name>
          .
          <year>2016</year>
          .
          <article-title>Toward a theory of variation in the organization of the word reading system</article-title>
          .
          <source>Scientific Studies of Reading</source>
          ,
          <volume>20</volume>
          (
          <issue>1</issue>
          ):
          <fpage>86</fpage>
          -
          <lpage>97</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref31">
        <mixed-citation>
          <string-name>
            <surname>Jenny R. Saffran</surname>
          </string-name>
          , Elissa L. Newport, Richard N. Aslin, Rachel A.
          <string-name>
            <surname>Tunick</surname>
            , and
            <given-names>Sandra</given-names>
          </string-name>
          <string-name>
            <surname>Barrueco</surname>
          </string-name>
          .
          <year>1997</year>
          .
          <article-title>Incidental language learning: Listening (and learning) out of the corner of your ear</article-title>
          .
          <source>Psychological science</source>
          ,
          <volume>8</volume>
          (
          <issue>2</issue>
          ):
          <fpage>101</fpage>
          -
          <lpage>105</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref32">
        <mixed-citation>
          <string-name>
            <given-names>Alastair C.</given-names>
            <surname>Smith</surname>
          </string-name>
          and
          <string-name>
            <given-names>Padraic</given-names>
            <surname>Monaghan</surname>
          </string-name>
          ,
          <year>2011</year>
          .
          <article-title>What are the functional units in reading? Evidence for statistical variation influencing word processing</article-title>
          , volume
          <volume>20</volume>
          , pages
          <fpage>159</fpage>
          -
          <lpage>172</lpage>
          . World Scientific.
          <source>Proceedings of the 12th Neural Computation and Psychology Workshop.</source>
        </mixed-citation>
      </ref>
      <ref id="ref33">
        <mixed-citation>
          <string-name>
            <given-names>Ross A.</given-names>
            <surname>Thompson</surname>
          </string-name>
          .
          <year>2009</year>
          .
          <article-title>Doing what doesn't come naturally</article-title>
          . Zero to Three Journal,
          <volume>30</volume>
          (
          <issue>2</issue>
          ):
          <fpage>33</fpage>
          -
          <lpage>39</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref34">
        <mixed-citation>
          <string-name>
            <given-names>Carole</given-names>
            <surname>Torgerson</surname>
          </string-name>
          , Greg Brooks, and
          <string-name>
            <given-names>Jill</given-names>
            <surname>Hall</surname>
          </string-name>
          .
          <year>2006</year>
          .
          <article-title>A systematic review of the research literature on the use of phonics in the teaching of reading and spelling</article-title>
          .
          <source>DfES Publications Nottingham.</source>
        </mixed-citation>
      </ref>
      <ref id="ref35">
        <mixed-citation>
          <string-name>
            <surname>Joseph K. Torgesen</surname>
          </string-name>
          , Richard K. Wagner, Carol A.
          <string-name>
            <surname>Rashotte</surname>
            , Elaine Rose, Patricia Lindamood, Tim Conway, and
            <given-names>Cyndi</given-names>
          </string-name>
          <string-name>
            <surname>Garvan</surname>
          </string-name>
          .
          <year>1999</year>
          .
          <article-title>Preventing reading failure in young children with phonological processing disabilities: Group and individual responses to instruction</article-title>
          .
          <source>Journal of Educational Psychology</source>
          ,
          <volume>91</volume>
          (
          <issue>4</issue>
          ):
          <fpage>579</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref36">
        <mixed-citation>
          <string-name>
            <given-names>Rebecca</given-names>
            <surname>Treiman</surname>
          </string-name>
          , John Mullennix, Ranka BijeljacBabic, and
          <string-name>
            <given-names>E.</given-names>
            <surname>Daylene</surname>
          </string-name>
          Richmond-Welty.
          <year>1995</year>
          .
          <article-title>The special role of rimes in the description, use, and acquisition of english orthography</article-title>
          .
          <source>Journal of Experimental Psychology: General</source>
          ,
          <volume>124</volume>
          (
          <issue>2</issue>
          ):
          <fpage>107</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref37">
        <mixed-citation>
          <string-name>
            <given-names>Rebecca</given-names>
            <surname>Treiman</surname>
          </string-name>
          , Brett Kessler,
          <string-name>
            <given-names>Jason D.</given-names>
            <surname>Zevin</surname>
          </string-name>
          , Suzanne Bick, and
          <string-name>
            <given-names>Melissa</given-names>
            <surname>Davis</surname>
          </string-name>
          .
          <year>2006</year>
          .
          <article-title>Influence of consonantal context on the reading of vowels: Evidence from children</article-title>
          .
          <source>Journal of Experimental Child Psychology</source>
          ,
          <volume>93</volume>
          (
          <issue>1</issue>
          ):
          <fpage>1</fpage>
          -
          <lpage>24</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref38">
        <mixed-citation>
          <string-name>
            <surname>Janet</surname>
            <given-names>I.</given-names>
          </string-name>
          <string-name>
            <surname>Vousden</surname>
          </string-name>
          .
          <year>2008</year>
          .
          <article-title>Units of english spelling-tosound mapping: a rational approach to reading instruction</article-title>
          .
          <source>Applied Cognitive Psychology: The Official Journal of the Society for Applied Research in Memory and Cognition</source>
          ,
          <volume>22</volume>
          (
          <issue>2</issue>
          ):
          <fpage>247</fpage>
          -
          <lpage>272</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref39">
        <mixed-citation>
          <string-name>
            <surname>Brendan S. Weekes</surname>
            ,
            <given-names>Anne E.</given-names>
          </string-name>
          <string-name>
            <surname>Castles</surname>
          </string-name>
          , and
          <string-name>
            <surname>Robert</surname>
            <given-names>A.</given-names>
          </string-name>
          <string-name>
            <surname>Davies</surname>
          </string-name>
          .
          <year>2006</year>
          .
          <article-title>Effects of consistency and age of acquisition on reading and spelling among developing readers</article-title>
          . Reading and Writing,
          <volume>19</volume>
          (
          <issue>2</issue>
          ):
          <fpage>133</fpage>
          -
          <lpage>169</lpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref40">
        <mixed-citation>
          <string-name>
            <given-names>Susan</given-names>
            <surname>Zeno</surname>
          </string-name>
          ,
          <string-name>
            <surname>Stephen H. Ivens</surname>
            , Robert T. Millard, and
            <given-names>Raj</given-names>
          </string-name>
          <string-name>
            <surname>Duvvuri</surname>
          </string-name>
          .
          <year>1995</year>
          .
          <article-title>The educator's word frequency guide</article-title>
          . Touchstone Applied Science Associates.
        </mixed-citation>
      </ref>
      <ref id="ref41">
        <mixed-citation>
          <string-name>
            <given-names>Johannes C.</given-names>
            <surname>Ziegler</surname>
          </string-name>
          and
          <string-name>
            <given-names>Usha</given-names>
            <surname>Goswami</surname>
          </string-name>
          .
          <year>2005</year>
          .
          <article-title>Reading acquisition, developmental dyslexia, and skilled reading across languages: a psycholinguistic grain size theory</article-title>
          .
          <source>Psychological bulletin</source>
          ,
          <volume>131</volume>
          (
          <issue>1</issue>
          ):
          <fpage>3</fpage>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>