<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article
  PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "http://dtd.nlm.nih.gov/publishing/3.0/journalpublishing3.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" article-type="research-article" dtd-version="3.0" xml:lang="en">
<front>
<journal-meta>
<journal-id journal-id-type="nlm-ta">PLoS ONE</journal-id>
<journal-id journal-id-type="publisher-id">plos</journal-id>
<journal-id journal-id-type="pmc">plosone</journal-id>
<journal-title-group>
<journal-title>PLOS ONE</journal-title>
</journal-title-group>
<issn pub-type="epub">1932-6203</issn>
<publisher>
<publisher-name>Public Library of Science</publisher-name>
<publisher-loc>San Francisco, CA USA</publisher-loc>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.1371/journal.pone.0123059</article-id>
<article-id pub-id-type="publisher-id">PONE-D-14-34743</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Research Article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Does Formal Complexity Reflect Cognitive Complexity? Investigating Aspects of the Chomsky Hierarchy in an Artificial Language Learning Study</article-title>
<alt-title alt-title-type="running-head">Testing the Chomsky Hierarchy</alt-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes" xlink:type="simple">
<name name-style="western">
<surname>Öttl</surname>
<given-names>Birgit</given-names>
</name>
<xref rid="aff001" ref-type="aff"><sup>1</sup></xref>
<xref rid="cor001" ref-type="corresp">*</xref>
</contrib>
<contrib contrib-type="author" xlink:type="simple">
<name name-style="western">
<surname>Jäger</surname>
<given-names>Gerhard</given-names>
</name>
<xref rid="aff002" ref-type="aff"><sup>2</sup></xref>
</contrib>
<contrib contrib-type="author" xlink:type="simple">
<name name-style="western">
<surname>Kaup</surname>
<given-names>Barbara</given-names>
</name>
<xref rid="aff001" ref-type="aff"><sup>1</sup></xref>
</contrib>
</contrib-group>
<aff id="aff001"><label>1</label> <addr-line>Department of Psychology, Eberhard Karls University, Tübingen, Germany</addr-line></aff>
<aff id="aff002"><label>2</label> <addr-line>Department of Linguistics, Eberhard Karls University, Tübingen, Germany</addr-line></aff>
<contrib-group>
<contrib contrib-type="editor" xlink:type="simple">
<name name-style="western">
<surname>Berent</surname>
<given-names>Iris</given-names>
</name>
<role>Academic Editor</role>
<xref ref-type="aff" rid="edit1"/>
</contrib>
</contrib-group>
<aff id="edit1"><addr-line>Northeastern University, UNITED STATES</addr-line></aff>
<author-notes>
<fn fn-type="conflict" id="coi001">
<p>The authors have declared that no competing interests exist.</p>
</fn>
<fn fn-type="con" id="contrib001">
<p>Conceived and designed the experiments: BO GJ BK. Performed the experiments: BO. Analyzed the data: BO. Wrote the paper: BO GJ BK.</p>
</fn>
<corresp id="cor001">* E-mail: <email xlink:type="simple">birgit.oettl@uni-tuebingen.de</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>17</day>
<month>4</month>
<year>2015</year>
</pub-date>
<pub-date pub-type="collection">
<year>2015</year>
</pub-date>
<volume>10</volume>
<issue>4</issue>
<elocation-id>e0123059</elocation-id>
<history>
<date date-type="received">
<day>5</day>
<month>10</month>
<year>2014</year>
</date>
<date date-type="accepted">
<day>12</day>
<month>2</month>
<year>2015</year>
</date>
</history>
<permissions>
<copyright-year>2015</copyright-year>
<copyright-holder>Öttl et al</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/" xlink:type="simple">
<license-p>This is an open access article distributed under the terms of the <ext-link ext-link-type="uri" xlink:href="http://creativecommons.org/licenses/by/4.0/" xlink:type="simple">Creative Commons Attribution License</ext-link>, which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited</license-p>
</license>
</permissions>
<self-uri content-type="pdf" xlink:href="info:doi/10.1371/journal.pone.0123059" xlink:type="simple"/>
<abstract>
<p>This study investigated whether formal complexity, as described by the Chomsky Hierarchy, corresponds to cognitive complexity during language learning. According to the Chomsky Hierarchy, nested dependencies (context-free) are less complex than cross-serial dependencies (mildly context-sensitive). In two artificial grammar learning (AGL) experiments participants were presented with a language containing either nested or cross-serial dependencies. A learning effect for both types of dependencies could be observed, but no difference between dependency types emerged. These behavioral findings do not seem to reflect complexity differences as described in the Chomsky Hierarchy. This study extends previous findings in demonstrating learning effects for nested and cross-serial dependencies with more natural stimulus materials in a classical AGL paradigm after only one hour of exposure. The current findings can be taken as a starting point for further exploring the degree to which the Chomsky Hierarchy reflects cognitive processes.</p>
</abstract>
<funding-group>
<funding-statement>While conducting this research the first author received support from the German National Academic Foundation. In addition, this research was supported by a grant from the German Research Foundation awarded to Barbara Kaup (SFB 833, Project B4) as well as by the ERC Advanced Grant 324246 Language Evolution: The Empirical Turn to Gerhard Jäger. The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript.</funding-statement>
</funding-group>
<counts>
<fig-count count="2"/>
<table-count count="2"/>
<page-count count="16"/>
</counts>
<custom-meta-group>
<custom-meta id="data-availability" xlink:type="simple">
<meta-name>Data Availability</meta-name>
<meta-value>Data have been deposited to an institutional repository: (<ext-link ext-link-type="uri" xlink:href="http://openscience.uni-leipzig.de/index.php/mr2/article/view/128" xlink:type="simple">http://openscience.uni-leipzig.de/index.php/mr2/article/view/128</ext-link>) (PID: 11022/0000-0000-46E0-4).</meta-value>
</custom-meta>
</custom-meta-group>
</article-meta>
</front>
<body>
<sec id="sec001" sec-type="intro">
<title>Introduction</title>
<sec id="sec002">
<title>Formal language theory and the Chomsky hierarchy</title>
<p>It has been a very fruitful guiding hypothesis of linguistic research since the mid-twentieth century that all natural languages are—despite their superficial diversity—fundamentally similar. While this general hypothesis is still controversial (see [<xref rid="pone.0123059.ref001" ref-type="bibr">1</xref>] for a skeptical view), it has led to many profound insights especially in the domain of grammar. To facilitate the study of the common core of natural language grammars with mathematical precision, Noam Chomsky (see for instance [<xref rid="pone.0123059.ref002" ref-type="bibr">2</xref>], [<xref rid="pone.0123059.ref003" ref-type="bibr">3</xref>]) proposed a number of idealizations, such as:
<list list-type="bullet">
<list-item><p>A natural language is considered as an infinite set of well-formed sentences, each of which is a finite string of words.</p></list-item>
<list-item><p>Whether or not a string of words is a grammatical sentence does not depend on its meaning.</p></list-item>
<list-item><p>The distinction between grammatical sentences and ungrammatical strings is categorical.</p></list-item>
</list></p>
<p>Every set of finite strings of symbols is a <italic>formal language</italic>. According to Chomsky, a comprehensive theory of syntax has to identify among the formal languages the <italic>possible human languages</italic>, i.e. the class of string sets that could be acquired as native language by a human infant. Chomsky [<xref rid="pone.0123059.ref004" ref-type="bibr">4</xref>] furthermore devised a classification of the formal languages into a nested hierarchy of complexity classes, the so-called <italic>Chomsky Hierarchy</italic> (a more comprehensive discussion of the Chomsky Hierarchy in relation to Artificial Grammar Learning can be found in [<xref rid="pone.0123059.ref005" ref-type="bibr">5</xref>]). The least restrictive—and therefore most complex—class are the <italic>Recursively enumerable</italic> or <italic>Type 0</italic> languages. These are all formal languages for which there is an algorithm enumerating all its elements. The more restrictive classes are the <italic>context-sensitive (Type 1)</italic> languages, the <italic>context-free (Type 2)</italic> languages, and the <italic>regular (Type 3)</italic> languages (the names <italic>context-sensitive</italic> and <italic>context-free</italic> are purely historically motivated and should not be taken at face value). It is uncontroversial among linguists that virtually all well-studied natural languages require at least context-free (the question whether this holds for all natural languages is currently hotly disputed; see [<xref rid="pone.0123059.ref006" ref-type="bibr">6</xref>], [<xref rid="pone.0123059.ref007" ref-type="bibr">7</xref>] and the references cited there) and not more than context-sensitive complexity. Whether or not all natural languages are context-free proved to be a fairly intricate problem which was only solved in 1984, when Huybregts [<xref rid="pone.0123059.ref008" ref-type="bibr">8</xref>] demonstrated that Swiss German is not context-free. Even Swiss German—and other natural languages that have non-context-free features—is much less complex than the most complex context-sensitive languages. Only a slight extension of the complexity of context-free languages is sufficient to cover all natural languages. Joshi and colleagues [<xref rid="pone.0123059.ref009" ref-type="bibr">9</xref>] proposed to refine the Chomsky Hierarchy by the additional level of <italic>mildly context-sensitive languages</italic> that include all context-free languages and are a proper sub-class of the context-sensitive languages. Based on current knowledge, all natural languages are mildly context-sensitive.</p>
<p>The three levels of the Chomsky Hierarchy that are relevant for the study of natural languages—regular, context-free and mildly context-sensitive languages—are characterized by the admissible <italic>dependencies</italic> within strings that they admit. Two positions within a string <italic>s</italic> that belongs to a language <italic>L</italic> are dependent if altering the symbol at the first position (such as deleting or replacing it, or adding additional material before or after) requires a concomitant change at the other position to preserve membership in L. To illustrate this with a simple example, consider the following English sentences:
<list list-type="order">
<list-item><p>
<list list-type="alpha-lower">
<list-item><p>It <italic>either</italic> rains <italic>or</italic> snows.</p></list-item>
<list-item><p>It <italic>neither</italic> rains <italic>nor</italic> snows.</p></list-item>
</list></p></list-item>
</list></p>
<p>Sentence (1a) is a grammatical sentence of English. If we replace <italic>either</italic> by <italic>neither</italic>, we also have to replace <italic>or</italic> by <italic>nor</italic> to preserve grammaticality. Therefore, there is a dependency between <italic>either</italic> and <italic>or</italic>.</p>
<p>Now let us consider a more complex pattern:
<list list-type="order">
<list-item><p>
<list list-type="alpha-lower">
<list-item><p>The cat<sub>1</sub> runs<sub>1</sub>.</p></list-item>
<list-item><p>The cat<sub>1</sub> that the dogs<sub>2</sub> know<sub>2</sub> runs<sub>1</sub>.</p></list-item>
<list-item><p>The cat<sub>1</sub> that the dogs<sub>2</sub> that the lady<sub>3</sub> owns<sub>3</sub> know<sub>2</sub> runs<sub>1</sub>.</p></list-item>
<list-item><p>The cat<sub>1</sub> that the dogs<sub>2</sub> that the lady<sub>3</sub> that … owns<sub>3</sub> know<sub>2</sub> runs<sub>1</sub>.</p></list-item>
</list></p></list-item>
</list></p>
<p>In English, the subject noun and the verb of a clause must agree in number—i.e. there is a dependency between the two positions—regardless of the number of words occurring between them. Such dependencies are called <italic>unbounded</italic>. In particular, we may insert an embedded clause between the two positions which contains its own subject and verb. This operation can be applied recursively, leading to an arbitrarily high number of nested dependencies (as far as the grammar of English is concerned, that is; the sentences quickly become incomprehensible due to processing constraints). Context-free languages, but not regular languages, may contain an unbounded number of nested dependencies. So the pattern above demonstrates English not to be regular. While context-free languages may contain an unbounded number of <italic>nested dependencies</italic>, they never contain an unbounded number of <italic>cross-serial dependencies</italic>. Mildly context-sensitive, but not context-free languages may contain this type of unbounded crossing dependencies. <xref rid="pone.0123059.g001" ref-type="fig">Fig 1</xref> shows an example string for nested and cross-serial dependencies and their location in the refined Chomsky Hierarchy.</p>
<fig id="pone.0123059.g001" position="float">
<object-id pub-id-type="doi">10.1371/journal.pone.0123059.g001</object-id>
<label>Fig 1</label>
<caption>
<title>The Chomsky Hierarchy including mildly context-sensitive languages.</title>
</caption>
<graphic mimetype="image" xlink:href="info:doi/10.1371/journal.pone.0123059.g001" position="float" xlink:type="simple"/>
</fig>
<p>As discussed above, the Chomsky Hierarchy classifies formal languages according to some quite abstract notion of complexity. It is far from obvious whether this notion of complexity corresponds to some empirically testable notion of cognitive complexity. Still, it has been hypothesized in the literature (most influentially in [<xref rid="pone.0123059.ref010" ref-type="bibr">10</xref>], [<xref rid="pone.0123059.ref011" ref-type="bibr">11</xref>]) that languages higher up in the hierarchy are harder to process—for humans as well as for other species—than those at the bottom of the hierarchy. This dovetails nicely with results from formal language theory regarding the processing complexity of these language classes. Time complexity of the recognition problem for regular languages is linear in length of the input string, while space complexity is constant [<xref rid="pone.0123059.ref012" ref-type="bibr">12</xref>]. In contradistinction, standard parsing algorithms for context-free languages (such as the CYK-algorithm) require cubic time complexity (meaning: the number of steps that a deterministic computer requires to decide whether a given string belongs to a given context-free grammar is bounded by a cubic function of the length of the string) and quadratic space complexity (meaning: the maximal number of memory cells is bounded by a quadratic function of the length of the string) [<xref rid="pone.0123059.ref013" ref-type="bibr">13</xref>]. Mildly context-free languages have a still higher processing complexity in this sense. Standard parsing algorithms for mildly context-sensitive languages (for Tree Adjoining Languages, to be precise [<xref rid="pone.0123059.ref014" ref-type="bibr">14</xref>]; the notion of “mild context-sensitivity” is sometimes also applied to a family of slightly more powerful language classes) such as the CYK-algorithm have a time complexity of O (<italic>n</italic><sup>6</sup>) and a space complexity of O (<italic>n</italic><sup>4</sup>) (meaning: the number of computing steps is bounded by a polynomial function of 6<sup>th</sup> degree of the length <italic>n</italic> of the string, and the number of memory cells by a polynomial function of 4<sup>th</sup> degree) [<xref rid="pone.0123059.ref013" ref-type="bibr">13</xref>]. This suggests the hypothesis that the cognitive processing complexity for humans of mildly context-sensitive languages is still higher than the complexity of context-free languages.</p>
<p>It should be added that the hypothesized correspondence between formal and cognitive processing complexity is, at best, suggestive. Leaving aside the obvious differences between deterministic Turing machines and the human brain, the mentioned results apply to general-purpose algorithms, i.e. algorithms that are capable of recognizing all regular/context-free/mildly context-sensitive languages. If human processing of sentences or of strings of some artificial languages employs more specialized strategies applicable only to sub-classes thereof, the mentioned complexity results do not necessarily carry over. Also, the complexity of the grammar induction process is orthogonal to the issue of processing complexity, and formal language theory has little to say about this. With these qualifications in mind, one possibility is to consider the Chomsky Hierarchy as a heuristics for processing complexity. This leads to the hypothesis that for humans, mildly context-sensitive languages are harder to process than context-free languages (which are in turn harder to process than regular languages). If so, cross-serial dependencies corresponding to the mildly-context sensitive complexity level in the Chomsky Hierarchy should be harder to process than nested dependencies and dependencies of the complexity level regular; nested dependencies should in turn be harder to process than less complex dependencies of the complexity level regular within the Chomsky Hierarchy. This hypothesis builds on the <italic>Derivational Theory of Complexity</italic> [<xref rid="pone.0123059.ref015" ref-type="bibr">15</xref>] according to which memory load and thus processing difficulty rises with increasing syntactic complexity [<xref rid="pone.0123059.ref016" ref-type="bibr">16</xref>]. However, it should be noted that in this theory syntactic complexity was defined as the number of transformations necessary to arrive at the deep structure of a sentence, and thus differed from the notion of syntactic complexity as defined by the Chomsky Hierarchy.</p>
</sec>
<sec id="sec003">
<title>Artificial grammar learning</title>
<p>The hypothesis that more complex languages are harder to process than less complex languages of course presupposes that human participants are in principle able to process cross-serial dependencies (and all less complex dependencies) even in languages that lack meaning or prosody, such as artificial languages. Studying syntactic dependencies in artificial languages has the advantage that syntactic processing can be studied in a very controlled and structured manner. In contrast to natural languages, language components such as semantics or prosody can be excluded from the cognitive process. Thus, syntactic processing can be studied in its pure form. Additionally, prior language knowledge which might vary between participants does not pose a problem.</p>
<p>An experimental paradigm that is widely used to study the cognitive processing of syntactic dependencies is the <italic>artificial grammar learning</italic> (AGL) paradigm [<xref rid="pone.0123059.ref017" ref-type="bibr">17</xref>]. In an AGL experiment, participants are presented with sequences of symbols following a particular rule. Here the idea is that symbols correspond to words, sequences correspond to sentences and the underlying rule corresponds to the syntactic dependencies between words in a sentence. Importantly, participants are blind to the rule underlying the sequences at the beginning of the experiment. After the first half of the experiment, participants are typically informed about the existence of the rule, and are asked to judge whether or not the subsequently presented sequences follow the rule from the first half of the experiment.</p>
<p>We will now turn to the results of studies investigating artificial grammar learning with languages of different complexity levels. Empirical evidence suggests that participants can easily process dependencies of the complexity type regular in artificial languages (i.e. [<xref rid="pone.0123059.ref010" ref-type="bibr">10</xref>], [<xref rid="pone.0123059.ref011" ref-type="bibr">11</xref>]). For nested dependencies, the empirical evidence is less clear because in studies using artificial languages it has been proven difficult to disentangle whether participants processed nested dependencies, as claimed by [<xref rid="pone.0123059.ref010" ref-type="bibr">10</xref>] and [<xref rid="pone.0123059.ref011" ref-type="bibr">11</xref>], or rather applied particular strategies not involving the processing of nested dependencies ([<xref rid="pone.0123059.ref018" ref-type="bibr">18</xref>], [<xref rid="pone.0123059.ref019" ref-type="bibr">19</xref>], [<xref rid="pone.0123059.ref020" ref-type="bibr">20</xref>]). More recent studies however demonstrated that participants are able to process nested and also cross-serial dependencies in artificial languages, at least under very restrictive conditions ([<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>], [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]) (we will come back to this point at the end of the introduction). It can be concluded that participants indeed seem to be able to process syntactic dependencies up to the complexity level mildly context-sensitive, in the absence of semantics, at least under very controlled and restricted conditions. Thus, the presupposition for investigating differences with respect to processing syntactic dependencies can be seen as satisfied.</p>
<p>Empirical studies assessing the difference in processing of nested and less complex regular dependencies provided evidence incorporable with the view that the Chomsky Hierarchy can be considered a heuristics for processing complexity ([<xref rid="pone.0123059.ref011" ref-type="bibr">11</xref>], [<xref rid="pone.0123059.ref023" ref-type="bibr">23</xref>], [<xref rid="pone.0123059.ref024" ref-type="bibr">24</xref>], for a review see [<xref rid="pone.0123059.ref025" ref-type="bibr">25</xref>]). With respect to the complexity difference between nested and cross-serial dependencies, Chesi and Moro [<xref rid="pone.0123059.ref026" ref-type="bibr">26</xref>] however argue against the idea of the Chomsky Hierarchy as an adequate reflection of cognitive language processes. Based on the <italic>Syntax Prediction Locality Theory</italic> (SPLT, [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>]), the authors propose that sentences with nested dependencies should be harder to process than sentences with cross-serial dependencies due to a higher memory load for nested dependencies compared to cross-serial dependencies [<xref rid="pone.0123059.ref026" ref-type="bibr">26</xref>]. Thus, the SPLT claims that the cognitive load involved in processing nested and cross-serial dependencies is not the same. More specifically, according to the SPLT [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>], when processing dependencies in a sentence, memory cost increases with increasing distance. That means that for example in <xref rid="pone.0123059.g001" ref-type="fig">Fig 1</xref>, the memory cost for the dependency between A<sub>1</sub> and B<sub>1</sub> in the nested sequence is higher (distance: four intervening elements) than the memory cost for the same dependency in the cross-serial sequence (distance: two intervening elements). When further comparing the memory costs for nested and cross-serial dependencies it becomes evident that for nested dependencies, distances between dependent elements differ between dependencies, whereas for cross-serial dependencies distances between dependent elements are the same for all dependencies. Because memory costs increase with increasing distance between dependent elements and long distance dependencies carry the most weight, nested dependencies are predicted to lead to higher memory costs compared to cross-serial dependencies. According to the SPLT this predicts more difficulties when processing nested dependencies compared to cross-serial dependencies (please refer to [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>] for a detailed derivation). Thus, [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>] and [<xref rid="pone.0123059.ref026" ref-type="bibr">26</xref>] suggest the opposite of what the Chomsky Hierarchy would predict. In line with this view, Bach and colleagues [<xref rid="pone.0123059.ref028" ref-type="bibr">28</xref>] showed that natural language sentences with nested dependencies are judged as less comprehensible compared to sentences with cross-serial dependencies.</p>
<p>Additional evidence for this view comes from two studies that investigated the processing of syntactic dependencies in artificial languages. Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] investigated the cognitive processing involved in cross-serial and nested dependencies. Participants were presented with visual sequences of letters containing either nested dependencies (context-free) or cross-serial dependencies (mildly context-sensitive) over a period of nine days. Letter sequences were embedded into a larger sequence of irrelevant letters to make the target sequences less obvious. Learning effects were observed for both types of dependencies, suggesting that humans are able to process dependencies of the mildly context-sensitive complexity level in an AGL paradigm when language learning lasts over several days. A processing advantage for cross-serial over nested dependencies also became evident in this study, consistent with [<xref rid="pone.0123059.ref026" ref-type="bibr">26</xref>] and the SPLT [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>] and contrary to what would be predicted by the Chomsky Hierarchy. However, it is important to note that the stimulus set employed in this study was fairly small. In addition, the stimuli were visual and not presented sequentially in this study. Thus, in order to further investigate the compatibility of the Chomsky Hierarchy with cognitive processes in an AGL setting it would certainly be beneficial to apply more natural learning conditions, such as a large stimulus set, auditory stimuli and a sequential presentation style. This was partly realized in a recent study by de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]. Participants were presented with auditory stimuli in sequential order. A combination of a serial reaction time (SRT) task and an AGL paradigm was applied. In addition, language exposure time was much shorter (approximately half an hour) as compared to Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>]. The results of this study were in line with the earlier findings in showing easier processing for cross-serial dependencies compared to nested dependencies. This study thus provides further evidence that cognitive processes for nested and cross-serial dependencies do not follow the predictions of the Chomsky Hierarchy when syntactic processing is investigated in an artificial language and thus independent of other language components. However, even though conditions in this study were closer to natural conditions, the applied set of stimuli was still fairly small. Thus, findings by [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] suggest that under very restricted experimental conditions participants are able to process nested and cross-serial dependencies and show higher performance for cross-serial compared to nested dependencies.</p>
<p>The present study aims at investigating whether the learning effect for both dependency types can be generalized to an experimental setting that is closer to natural conditions, in particular concerning the set size of the stimulus material. In addition, the present study investigates whether under these conditions, performance for nested dependencies is better than that for cross-serial dependencies (as predicted by the Chomsky Hierarchy), or whether the differences are the other way around (as predicted by SPLT). As in the study by de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] auditory stimuli were presented sequentially to the participants and language exposure time was rather short. An AGL paradigm was applied comparable to the study by Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>]. By converging the experimental settings of Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>], the present study allows investigating learning effects without extensive language exposure and when participants are confronted with language material that is closer to natural language. In line with findings from de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] and Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] we expect to find learning effects for both nested and cross-serial dependencies, but with better performance for the cross-serial compared to nested dependencies.</p>
</sec>
</sec>
<sec id="sec004">
<title>Experiment 1</title>
<sec id="sec005" sec-type="materials|methods">
<title>Method</title>
<sec id="sec006">
<title>Participants</title>
<p>Thirty participants took part in the experiment (age: <italic>M</italic>(<italic>SD</italic>) = 26.87 (4.02) years; gender: 24 female; native language German: 29). All participants were right handed and received course credit or a financial reimbursement of 8 Euro per hour for participating in the study. Participants were randomly assigned to the nested dependency group or the cross-serial dependency group (15 participants per group).</p>
</sec>
<sec id="sec007">
<title>Ethics statement</title>
<p>The experimental testing was in agreement with the guidelines for good scientific practice at the University of Tübingen (Germany). This was checked and approved by the Head of Psychology, Faculty of Science, University of Tübingen. He functioned as an independent individual judge who was in no way involved in the study. Prior to the experiment participants were informed that they were free to terminate the experiment at any time without facing disadvantages. After participants signed an informed consent form, a number was assigned to each participant. This number was associated with the recorded data of each participant throughout the whole experiment and data analysis. Thus, participants' anonymity was always preserved; at no point could the recorded data be associated with a participant's name.</p>
</sec>
<sec id="sec008">
<title>Apparatus and Stimuli</title>
<p>Spoken auditory stimuli consisted of 20 syllable pairs in which one syllable belonged to category A and one syllable to category B, see <xref rid="pone.0123059.t001" ref-type="table">Table 1</xref>. As in the study by Friederici and colleagues [<xref rid="pone.0123059.ref011" ref-type="bibr">11</xref>], category membership was indicated by the vowel of each syllable (category A: ‘e’, ‘i’, category B: ‘o’, ‘u’). Element pairing was signified by the identical first letter of both syllables within a pair. In the example syllable pair ‘del—dol’, ‘del’ belongs to category A and ‘dol’ to category B. The first letter ‘d’ indicates that both syllables belong to the same element pair. All syllables were spoken by a female native German speaker and were recorded with Audacity (syllable length <italic>M</italic>(<italic>range</italic>) = 689.95 ms (476 ms–906 ms)). Syllables were merged to sequences of either 4 (short), 6 (medium) or 8 (long) syllables using Matlab (R2011b, 32-bit win). As a result of this, there was no rising or falling intonation across the whole sequence. All syllables were counterbalanced across syllable position and number of occurrence for the Category A elements (first half of sequence). Category B elements (second half of sequence) were separately re-arranged for each sequence length according to either nested dependencies (reverse order of A elements) or cross-serial dependencies (same order as A elements). As a consequence, there were minor differences for the nested dependencies with respect to how often a syllable occurred in each position in the B part of the sequences. All sequences were presented only once throughout the entire experiment in one random order to all participants. During the learning phase, participants were presented with 240 sequences (80 sequences each for short, medium, long) following one language (containing either nested dependencies or cross-serial dependencies). In the test phase, participants were then presented with 360 sequences, of which 180 sequences (60 sequences for short, medium, long) belonged to the dependency type in the learning phase (correct trials) and 180 sequences (60 sequences each for short, medium, long) did not belong to the dependency type in the learning phase (incorrect trials). The sequences presented in incorrect trials always followed the language that was not presented in the learning phase. The experiment was programed in Matlab (R2010a, 32-bit maci) using the Psychophysics Toolbox (Version 3.0.8). Participants performed the experiment on a MacBook Pro and were presented with the auditory stimuli via headphones. A short questionnaire was completed after the experiment to obtain information about potential strategy usage.</p>
<table-wrap id="pone.0123059.t001" position="float">
<object-id pub-id-type="doi">10.1371/journal.pone.0123059.t001</object-id>
<label>Table 1</label> <caption><title>Stimulus material of Experiment 1.</title></caption>
<alternatives>
<graphic id="pone.0123059.t001g" position="float" mimetype="image" xlink:href="info:doi/10.1371/journal.pone.0123059.t001" xlink:type="simple"/>
<table>
<colgroup span="1">
<col align="left" valign="middle" span="1"/>
<col align="left" valign="middle" span="1"/>
</colgroup>
<thead>
<tr>
<th align="left" rowspan="1" colspan="1">Class A</th>
<th align="left" rowspan="1" colspan="1">Class B</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" rowspan="1" colspan="1">del</td>
<td align="left" rowspan="1" colspan="1">dol</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">sted</td>
<td align="left" rowspan="1" colspan="1">stod</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">bem</td>
<td align="left" rowspan="1" colspan="1">bom</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">jelz</td>
<td align="left" rowspan="1" colspan="1">jolz</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">pfes</td>
<td align="left" rowspan="1" colspan="1">pfos</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">schip</td>
<td align="left" rowspan="1" colspan="1">schop</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">riw</td>
<td align="left" rowspan="1" colspan="1">row</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">fid</td>
<td align="left" rowspan="1" colspan="1">fod</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">hiz</td>
<td align="left" rowspan="1" colspan="1">hoz</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">zib</td>
<td align="left" rowspan="1" colspan="1">zob</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">lef</td>
<td align="left" rowspan="1" colspan="1">luf</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">sek</td>
<td align="left" rowspan="1" colspan="1">suk</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">kem</td>
<td align="left" rowspan="1" colspan="1">kum</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">pegs</td>
<td align="left" rowspan="1" colspan="1">pugs</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">wel</td>
<td align="left" rowspan="1" colspan="1">wul</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">tix</td>
<td align="left" rowspan="1" colspan="1">tux</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">miv</td>
<td align="left" rowspan="1" colspan="1">muv</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">nist</td>
<td align="left" rowspan="1" colspan="1">nust</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">xim</td>
<td align="left" rowspan="1" colspan="1">xum</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">gid</td>
<td align="left" rowspan="1" colspan="1">gud</td>
</tr>
</tbody>
</table>
</alternatives>
</table-wrap>
</sec>
<sec id="sec009">
<title>Procedure</title>
<p>Throughout the entire experiment, a white fixation cross appeared on a grey background while a sequence was played. After every twentieth trial, participants had the option to take a short break. In the learning phase, a sequence could be initiated by pressing the space bar, and participants were instructed to listen to the syllable sequences. Importantly, they were not informed about the existence of a rule underlying the sequences. In 25% of the trials, they were asked to repeat the last heard sequence in order to assess their state of alertness. The test phase followed immediately after the learning phase. Participants were informed that all sequences from the previous phase followed an underlying rule. They were further informed that they would now be presented with sequences either consistent with the rule from the learning phase or not. They were instructed to judge for each sequence whether it followed the rule, by pressing the c-key for “yes” and the m-key for “no”. Participants could take as much time as needed to indicate their decision by button press. The experiment lasted approximately one hour.</p>
</sec>
</sec>
<sec id="sec010" sec-type="results">
<title>Results</title>
<p>All reported analyses were performed in Matlab (R2011b, 32-bit win) and SPSS (version 20). D prime (<italic>d’</italic>) as a measure for performance was calculated for each dependency type. In order to ensure that none of the four classes hits, misses, false alarms and correct rejections equaled zero, 0.5 was added to each class for all participants [<xref rid="pone.0123059.ref029" ref-type="bibr">29</xref>]. A <italic>d’</italic> of 0 corresponds to chance level (50%). Greenhouse-Geisser correction for sphericity violation in repeated measures ANOVAs was applied when appropriate. In case of a significant finding, effect size was calculated using Cohen’s d (<italic>d</italic>) for <italic>t</italic>-tests and partial eta square (<italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup><italic>)</italic> for ANOVAs.</p>
<p>To evaluate the performance of the participants, a <italic>t</italic>-test against zero was calculated for each dependency type. In both groups, participants performed significantly above chance (nested: <italic>t</italic>(14) = 4.45, <italic>p</italic> = .001, <italic>d</italic> = 1.15; cross-serial: <italic>t</italic>(14) = 5.55, <italic>p</italic> &lt;.001, <italic>d</italic> = 1.43). Thus, learning effects were present for both types of dependency. There was no difference between the two nested and cross-serial dependencies with respect to performance as indicated by the results of a <italic>t</italic>-test for independent samples, (<italic>t</italic>(28) = 0.08; <italic>p</italic> &gt;.90).</p>
<p>To assess potential performance differences between the two dependencies at earlier time points in the test phase, we divided the test phase into 3 blocks (Block 1: trial 1–120, Block 2: trial 121–240, Block 3: trial 241–360) and calculated <italic>d’</italic> for each block separately. A 2-x-3 ANOVA with the factors dependency (nested, cross-serial) and block (1, 2, 3) revealed a significant main effect of block (<italic>F</italic>(1.30,36.46) = 5.51, <italic>p</italic> &lt;.05, <italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup> = 0.16), but no interaction (<italic>F</italic> &lt;1) and no main effect of dependency (<italic>F</italic> &lt;1). Follow-up 2-x-2 ANOVAs showed a significant improvement from Block 1 to Block 2 (<italic>F</italic>(1,28) = 9.57, <italic>p</italic> &lt;.01, <italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup> = 0.26) and from Block 1 to Block 3 (<italic>F</italic>(1,28) = 4.51, <italic>p</italic> &lt;.05, <italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup> = 0.14) but not from Block 2 to Block 3 (<italic>F</italic> &lt;1).</p>
<p>To investigate performance for sequences of different lengths, <italic>d’</italic> was calculated for each sequence length separately. A 2-x-3 ANOVA with the factors dependency (nested, cross-serial) and sequence length (short, medium, long) showed a significant main effect of sequence length (<italic>F</italic>(1.41,39.45) = 27.88, <italic>p</italic> &lt;.001, <italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup> = 0.50) but no main effect of language (<italic>F</italic>&lt;1) and no interaction (<italic>F</italic>(1.41,39.45) = 2.07, <italic>p</italic> = .15). Separate 2-x-2 ANOVAs with the factors dependency and sequence length indicated that performance was better for short sequences as compared to medium (<italic>F</italic>(1,28) = 43.74, <italic>p</italic> &lt;.001, <italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup> = 0.61) and long sequences (<italic>F</italic>(1,28) = 26.83, <italic>p</italic> &lt;.001, <italic>η</italic><sub><italic>p</italic></sub><sup><italic>2</italic></sup> = 0.49). No difference in performance was observed for medium and long sequences (<italic>F</italic> &lt;1).</p>
</sec>
<sec id="sec011" sec-type="conclusions">
<title>Discussion</title>
<p>Results showed that for both types of dependencies, a learning effect was observed but performance did not differ between nested and cross-serial dependencies. Throughout the test phase performance seemed to improve for both types of dependency in a similar way. This suggests that learning effects for nested and cross-serial dependencies in an AGL paradigm are also present under more natural conditions (larger stimulus set, spoken auditory stimuli, sequential presentation style). Furthermore, as in previous studies, no evidence could be obtained for the idea that the Chomsky Hierarchy reflects cognitive processes, at least not with regard to the proposed complexity difference for nested and cross-serial dependencies.</p>
<p>However, our results are not fully consistent with these earlier findings by de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] and Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] in showing no better performance for cross-serial as compared to the nested dependencies. Participants in our study showed a higher learning performance for short compared to medium or long sequences, which is consistent with the results of previous studies showing that processing difficulty rises for materials with more than two dependencies ([<xref rid="pone.0123059.ref028" ref-type="bibr">28</xref>], [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]).</p>
<p>Interestingly, two participants in the nested dependencies group reported that they rated the sequences based on the first element pair (i.e.: nested dependencies: <bold>A</bold><sub><bold>1</bold></sub>A<sub>2</sub>A<sub>3</sub>B<sub>3</sub>B<sub>2</sub><bold>B</bold><sub><bold>1</bold></sub>; cross-serial dependencies: <bold>A</bold><sub><bold>1</bold></sub>A<sub>2</sub>A<sub>3</sub><bold>B</bold><sub><bold>1</bold></sub>B<sub>2</sub>B<sub>3</sub>). As incorrect sequences in the test phase were sequences from the respective other dependency, sequences necessarily differed in their arrangement of the element pairs. Hence, by paying attention to only one element pair and ignoring the rest of the sequence, it was possible to respond correctly in the test phase without having learnt the underlying language. Importantly, this alternative strategy would not reflect language processing on a context-free or mildly context-sensitive complexity level but rather the acquisition of a dependency type of the less complex regular complexity level. Attending only to the first element pair will be referred to as <italic>first-element-pair-strategy</italic> in what follows. Tracking only the last element pair while ignoring the rest of the sequence will be referred to as <italic>last-element-pair-strategy</italic>. It can be concluded that the application of one of these strategies might have affected learning performance in this experiment and possibly overshadowed performance differences between the two dependency types. Therefore, Experiment 2 was conducted to assess the degree to which learning effects from Experiment 1 can be explained by these strategies. The learning phase was kept identical to Experiment 1. To control for the application of alternative strategies, incorrect sequences in the test phase were adapted.</p>
</sec>
</sec>
<sec id="sec012">
<title>Experiment 2</title>
<sec id="sec013" sec-type="materials|methods">
<title>Method</title>
<sec id="sec014">
<title>Participants</title>
<p>Forty-four participants took part in the experiment. Four participants had to be excluded from the analyses because it could not be guaranteed that they were blind with respect to the underlying rule at the beginning of the experiment. These four participants were most likely not blind to the underlying rule because they already had taken part in an AGL study in our laboratory in the past (two participants), reported after the experiment that they had been told by a friend that the goal of the learning phase was to find a rule prior to the experiment (one participant), completed the experiment while the instruction sheet for the experimenter (including the goal to find the rule in the learning phase) was accidently left in the cabin with the participant (one participant). Out of the 40 remaining participants (age: <italic>M</italic>(<italic>SD</italic>) = 21.78(3.83) years; gender: 32 female; native language German: 40) eight participants were left handed. Participants received course credit or a financial reimbursement of 8 Euro per hour. As in Experiment 1, participants were randomly assigned to one of the two learning groups (20 participants per group).</p>
</sec>
<sec id="sec015">
<title>Ethics statement</title>
<p>The experimental testing was in agreement with the guidelines for good scientific practice at the University of Tübingen (Germany). This was checked and approved by the Head of Psychology, Faculty of Science, University of Tübingen. He functioned as an independent individual judge who was in no way involved in the study. Prior to the experiment participants were informed that they were free to terminate the experiment at any time without facing disadvantages. After participants signed an informed consent form, a number was assigned to each participant. This number was associated with the recorded data of each participant throughout the whole experiment and data analysis. Thus, participants' anonymity was always preserved; at no point could the recorded data be associated with a participant's name.</p>
</sec>
<sec id="sec016">
<title>Apparatus and Stimuli</title>
<p>The stimulus material was identical to the stimulus material from Experiment 1 except for the syllable sequences in the test phase. In the test phase of this experiment, participants were only presented with sequences of medium length (n = 240). Half of the trials were sequences consistent with the dependency type from the previous learning phase (correct trials) and the other half of trials were inconsistent with the dependency type from the learning phase (incorrect trials) (see <xref rid="pone.0123059.t002" ref-type="table">Table 2</xref>). One third of the incorrect trials were sequences following the dependency type not presented in the learning phase (Subset 1; similar to Experiment 1). In order to control for the first-element-pair-strategy, one third of the incorrect trials were sequences violating the dependency type from the learning phase (target rule) while preserving the first element pair (Subset 2; see 3rd and 4th incorrect trial of nested dependency and cross-serial dependency in <xref rid="pone.0123059.t002" ref-type="table">Table 2</xref>). Thus, participants only tracking the first element pair would judge this sequence as correct even though it does not follow the target rule. To control for the last-element-pair-strategy, the remaining third of the incorrect trials violated the target rule while preserving the last element pair in the sequence (Subset 3; see 5th and 6th incorrect trial of nested dependencies and cross-serial dependencies in <xref rid="pone.0123059.t002" ref-type="table">Table 2</xref>). Thus, participants who are only paying attention to the third (last) element pair would incorrectly judge these sequences as correct.</p>
<table-wrap id="pone.0123059.t002" position="float">
<object-id pub-id-type="doi">10.1371/journal.pone.0123059.t002</object-id>
<label>Table 2</label> <caption><title>Correct and incorrect trials in Experiment 2.</title></caption>
<alternatives>
<graphic id="pone.0123059.t002g" position="float" mimetype="image" xlink:href="info:doi/10.1371/journal.pone.0123059.t002" xlink:type="simple"/>
<table>
<colgroup span="1">
<col align="left" valign="middle" span="1"/>
<col align="left" valign="middle" span="1"/>
<col align="left" valign="middle" span="1"/>
<col align="left" valign="middle" span="1"/>
</colgroup>
<thead>
<tr>
<th align="left" rowspan="1" colspan="1">Nested dependencies</th>
<th align="left" rowspan="1" colspan="1">Cross-serial dependencies</th>
<th align="left" rowspan="1" colspan="1">Accuracy</th>
<th align="left" rowspan="1" colspan="1">Subset and Strategy</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" rowspan="1" colspan="1">A<sub>1</sub>A<sub>2</sub>A<sub>3</sub>B<sub>3</sub>B<sub>2</sub>B<sub>1</sub></td>
<td align="left" rowspan="1" colspan="1">A<sub>1</sub>A<sub>2</sub>A<sub>3</sub>B<sub>1</sub>B<sub>2</sub>B<sub>3</sub></td>
<td align="left" rowspan="1" colspan="1">Correct</td>
<td align="left" rowspan="1" colspan="1"/>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">A<sub>1</sub><bold>A</bold><sub><bold>2</bold></sub>A<sub>3</sub>B<sub>1</sub><bold>B</bold><sub><bold>2</bold></sub>B<sub>3</sub></td>
<td align="left" rowspan="1" colspan="1">A<sub>1</sub><bold>A</bold><sub><bold>2</bold></sub>A<sub>3</sub>B<sub>3</sub><bold>B</bold><sub><bold>2</bold></sub>B<sub>1</sub></td>
<td align="left" rowspan="1" colspan="1">Incorrect</td>
<td align="left" rowspan="1" colspan="1">Subset 1: Similar to Experiment 1</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">A<sub>3</sub><bold>A</bold><sub><bold>2</bold></sub>A<sub>1</sub>B<sub>3</sub><bold>B</bold><sub><bold>2</bold></sub>B<sub>1</sub></td>
<td align="left" rowspan="1" colspan="1">A<sub>3</sub><bold>A</bold><sub><bold>2</bold></sub>A<sub>1</sub>B<sub>1</sub><bold>B</bold><sub><bold>2</bold></sub>B<sub>3</sub></td>
<td align="left" rowspan="1" colspan="1">Incorrect</td>
<td align="left" rowspan="1" colspan="1">Subset 1: Similar to Experiment 1</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1"><bold>A</bold><sub><bold>1</bold></sub>A<sub>2</sub>A<sub>3</sub>B<sub>2</sub>B<sub>3</sub><bold>B</bold><sub><bold>1</bold></sub></td>
<td align="left" rowspan="1" colspan="1"><bold>A</bold><sub><bold>1</bold></sub>A<sub>2</sub>A<sub>3</sub><bold>B</bold><sub><bold>1</bold></sub>B<sub>3</sub>B<sub>2</sub></td>
<td align="left" rowspan="1" colspan="1">Incorrect</td>
<td align="left" rowspan="1" colspan="1">Subset 2: First-element-pair-strategy</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1"><bold>A</bold><sub><bold>1</bold></sub>A<sub>3</sub>A<sub>2</sub>B<sub>3</sub>B<sub>2</sub><bold>B</bold><sub><bold>1</bold></sub></td>
<td align="left" rowspan="1" colspan="1"><bold>A</bold><sub><bold>1</bold></sub>A<sub>3</sub>A<sub>2</sub><bold>B</bold><sub><bold>1</bold></sub>B<sub>2</sub>B<sub>3</sub></td>
<td align="left" rowspan="1" colspan="1">Incorrect</td>
<td align="left" rowspan="1" colspan="1">Subset 2: First-element-pair-strategy</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">A<sub>1</sub>A<sub>2</sub><bold>A</bold><sub><bold>3</bold></sub><bold>B</bold><sub><bold>3</bold></sub>B<sub>1</sub>B<sub>2</sub></td>
<td align="left" rowspan="1" colspan="1">A<sub>1</sub>A<sub>2</sub><bold>A</bold><sub><bold>3</bold></sub>B<sub>2</sub>B<sub>1</sub><bold>B</bold><sub><bold>3</bold></sub></td>
<td align="left" rowspan="1" colspan="1">Incorrect</td>
<td align="left" rowspan="1" colspan="1">Subset 3: Last-element-pair-strategy</td>
</tr>
<tr>
<td align="left" rowspan="1" colspan="1">A<sub>2</sub>A<sub>1</sub><bold>A</bold><sub><bold>3</bold></sub><bold>B</bold><sub><bold>3</bold></sub>B<sub>2</sub>B<sub>1</sub></td>
<td align="left" rowspan="1" colspan="1">A<sub>2</sub>A<sub>1</sub><bold>A</bold><sub><bold>3</bold></sub>B<sub>1</sub>B<sub>2</sub><bold>B</bold><sub><bold>3</bold></sub></td>
<td align="left" rowspan="1" colspan="1">Incorrect</td>
<td align="left" rowspan="1" colspan="1">Subset 3: Last-element-pair-strategy</td>
</tr>
</tbody>
</table>
</alternatives>
<table-wrap-foot>
<fn id="t002fn001"><p><italic>Note</italic>. For incorrect trials element pairs consistent with the respective language are bold faced.</p></fn>
</table-wrap-foot>
</table-wrap>
</sec>
<sec id="sec017">
<title>Procedure</title>
<p>The procedure was identical to the procedure in Experiment 1 with the only difference being that each participant was now presented with a different random order of the stimuli.</p>
</sec>
</sec>
<sec id="sec018" sec-type="results">
<title>Results</title>
<p>Experiment 2 investigated the learnability of nested dependencies and cross-serial dependencies while controlling for the first- and the last-element-pair-strategy. Data analysis was performed as in Experiment 1.</p>
<p><xref rid="pone.0123059.g002" ref-type="fig">Fig 2</xref> presents the overall performance for each group in percentage correct along with the corresponding <italic>d’</italic>. Results showed that participants performed significantly above chance for both types of dependencies (nested: <italic>t</italic>(19) = 3.65, <italic>p</italic> &lt;.01, <italic>d</italic> = 0.82; cross-serial: <italic>t</italic>(19) = 4.33, <italic>p &lt;</italic>.001, <italic>d</italic> = 0.97). Thus, as in Experiment 1, a learning effect was observed for both languages. Similar to Experiment 1, no differences in performance between dependency types was observed (<italic>t</italic>(38) = -0.56; <italic>p</italic> = .58) leading to the conclusion that both dependencies were learned equally well.</p>
<fig id="pone.0123059.g002" position="float">
<object-id pub-id-type="doi">10.1371/journal.pone.0123059.g002</object-id>
<label>Fig 2</label>
<caption>
<title>Results from Experiment 2.</title>
<p>A: Mean percentage correct <italic>(SE)</italic> for nested and cross-serial dependencies, respectively. B: Mean <italic>d’s (SE)</italic> for nested and cross-serial dependencies, respectively.</p>
</caption>
<graphic mimetype="image" xlink:href="info:doi/10.1371/journal.pone.0123059.g002" position="float" xlink:type="simple"/>
</fig>
<p>In order to control for the use of the first- and the last-element-strategy, <italic>d’</italic> was calculated with an error rate taking only sequences for the first- or the last-element-strategy into account (‘<italic>d’</italic><sub><italic>first</italic></sub>’, ‘<italic>d’</italic><sub><italic>last</italic></sub>’). The idea was that participants applying one of these strategies would rate incorrect sequences that preserve the particular element pair falsely as correct. Thus, taking only these strategy-specific sequences into account, a higher error rate and hence a smaller <italic>d’</italic> should be observed for strategic response behavior. In other words, a low <italic>d’</italic><sub><italic>first</italic></sub> and a low <italic>d’</italic><sub><italic>last</italic></sub> indicate that participants applied the respective strategy, whereas a high <italic>d’</italic><sub><italic>first</italic></sub> and a high <italic>d’</italic><sub><italic>last</italic></sub> show high performance independent of strategic behavior. For both dependency types, <italic>d’</italic><sub><italic>first</italic></sub> was significantly above chance (nested: <italic>t</italic>(19) = 3.06, <italic>p</italic> &lt;.01, <italic>d</italic> = 0.68; cross-serial: <italic>t</italic>(19) = 4.35, <italic>p</italic> &lt;.001, <italic>d</italic> = 0.97) as well as <italic>d’</italic><sub><italic>last</italic></sub> (nested: <italic>t</italic>(19) = 3.37, <italic>p</italic> &lt;.01, <italic>d</italic> = 0.75; cross-serial: <italic>t</italic>(19) = 3.34, <italic>p</italic> &lt;.01, <italic>d</italic> = 0.75). Thus, participants showed learning effects for both types of dependencies independent of the first-element-pair-strategy or the last-element-pair-strategy.</p>
<p>Whether or not participants applied the first-element-pair-strategy or the last-element-pair-strategy was additionally assessed on an individual participants’ analysis. Here we compared the number of yes- to the number of no-responses in each subset of the incorrect sequences. If participants did not learn the underlying dependency, they should press ‘yes’ as often as ‘no’ in Subset 2 and Subset 3 of incorrect sequences. On the other hand, if participants acquired the underlying type of dependency, they should press ‘no’ more often than ‘yes’ in Subset 2 and Subset 3 of incorrect stimuli. In contrast, if participants applied a strategy they should press ‘yes’ more often than ‘no’ in Subset 2 and Subset 3 of the incorrect stimuli. To assess whether participants generally tended to press ‘yes’ more often than ‘no’ independent of any rule knowledge or strategy usage, yes/no- responses were also analyzed in Subset 1 of the incorrect stimuli. In this subset, sequences neither followed the first-element-pair-strategy nor the last-element-pair-strategy. Three binomial tests with a critical value of 0.5 (corresponding to chance) were calculated for each participant separately (first-element-pair strategy, last-element-pair strategy, general tendency to say ‘yes’). Out of all participants, three participants in the nested dependency group responded significantly more often with ‘yes’ in one of the two strategy subsets (Subset 2 and Subset 3), but importantly did not respond significantly more often with ‘yes’ in the general tendency subset (Subset 1). When these three participants were left out of the main analysis, learning performance within and between groups was in line with the results reported above. We therefore consider it safe to conclude that the observed learning effects do not reflect the use of the first- or the last-element-pair strategies.</p>
<p>As in Experiment 1, performance was assessed throughout the test phase at three time points (Block 1: trial 1–80, Block 2: trial 81–160, Block 3: trial 161–240). A 2-x-3 ANOVA with the factors dependency type (nested, cross-serial) and block (1, 2, 3) revealed no significant main effect of block (<italic>F</italic>(1.58,60.13) = 2.31, <italic>p</italic> = .12), no significant main effect of dependency type (<italic>F</italic> &lt;1) and no significant interaction (<italic>F</italic>(1.58,60.13) = 2.45, <italic>p</italic> = .11). Thus, performance for both types of dependency did not differ significantly across the test phase.</p>
</sec>
<sec id="sec019" sec-type="conclusions">
<title>Discussion</title>
<p>Results showed that when controlling for the first-element-pair and the last-element-pair-strategy, participants still showed a learning effect for both types of dependency. Thus, Experiment 2 replicated the finding from Experiment 1 while controlling for strategic behavior.</p>
</sec>
</sec>
<sec id="sec020">
<title>General Discussion</title>
<p>This study investigated whether the Chomsky Hierarchy is reflected in cognitive learning processes when participants are confronted with new (artificial) language material. Nested dependencies (context-free) and cross-serial dependencies (mildly context-sensitive) were investigated. According to the Chomsky Hierarchy, cross-serial dependencies are located at a higher level of complexity than nested dependencies, which should lead to lower learning performance for cross-serial compared to nested dependencies in a cognitive learning task. In contrast, based on [<xref rid="pone.0123059.ref028" ref-type="bibr">28</xref>], the SPLT [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>] would predict a higher memory load when processing nested compared to cross-serial dependencies in natural language. Therefore, according to the SPLT, processing nested dependencies should result in more difficulties than processing cross-serial dependencies in natural language, a prediction that is also made by Chesi and Moro [<xref rid="pone.0123059.ref026" ref-type="bibr">26</xref>]. Furthermore, de Vries and colleagues [<xref rid="pone.0123059.ref030" ref-type="bibr">30</xref>] proposed that nested dependencies involving a reversed copying process should be more difficult to learn than the cross-serial dependencies involving a simple copying process in artificial languages. Empirical studies provided evidence for this hypothesis in showing better learning performance for cross-serial as compared to nested dependencies in AGL experiments after extensive language exposure [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and a processing benefit for cross-serial over the nested dependencies in an SRT-AGL experiment after approximately half an hour [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]. Adding onto these findings, the current study aimed at investigating the cognitive adequacy of the Chomsky Hierarchy by combining the approaches by de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] and Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>]. In an AGL paradigm [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] spoken auditory stimuli were presented in a sequential presentation style [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]. Language exposure was rather short consistent with de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]. Importantly, more natural stimulus conditions were realized by constructing a larger set of stimuli. We hypothesized that participants would acquire both types of dependencies in our experimental setup. In line with Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] we expected to find learning effects for both dependency types but better performance for the cross-serial than for nested dependencies.</p>
<p>Results showed a learning effect for both types of dependencies, but no difference between nested and cross-serial dependencies in Experiment 1. In Experiment 1, we could not be sure that participants did indeed learn the underlying dependency without applying the first-element-pair-strategy and/or the last-element-pair-strategy. Therefore, Experiment 2 was conducted, controlling for the use of alternative strategies. Results from Experiment 2 were in line with findings from Experiment 1, suggesting that nested dependencies, reflecting the context-free complexity level, and cross-serial dependencies, reflecting the mildly context-sensitive complexity level, are learnable in principle with our artificial language material. These results therefore extent existing findings [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] by demonstrating that participants are able to acquire nested and cross-serial dependencies within only one hour of exposure in a larger set of artificial language stimuli. Beyond this, however, our results do not match the predictions that can be derived from the Chomsky Hierarchy, and are in this respect consistent with the findings by Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and de Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>].</p>
<p>The absence of a significant performance difference between nested and cross-serial dependencies in our study is however also not in line with predictions by the SPLT [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>] predicting better performance for cross-serial compared to nested dependencies for natural language processing. This might suggest that predictions by the SPLT for language comprehension [<xref rid="pone.0123059.ref027" ref-type="bibr">27</xref>] are not easily transferable to the processing of artificial languages where syntactic processing is studied in isolation. Findings from [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] and [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] that investigated the processing of nested and cross-serial dependencies in artificial languages however speak against this conclusion since their findings are incorporable with predictions by the SPLT.</p>
<p>One explanation for why performance for cross-serial dependencies did not differ from the performance for nested dependencies in our study could be that nested dependencies might have had a processing advantage relative to cross-serial dependencies. More specifically, it is possible that processing the innermost dependency of nested dependency sequences (e.g. ‘A<sub>3</sub> B<sub>3</sub>’ in the sequence A<sub>1</sub> A<sub>2</sub> A<sub>3</sub> B<sub>3</sub> B<sub>2</sub> B<sub>1</sub>) was particularly easy for participants in our study because no element intervened the dependent elements at this position. It would be interesting to investigate whether this potential processing advantage for nested dependencies could be demolished by adding an intervening element at the innermost position for both dependency types, i.e. a “dummy” syllable (‘D’) (we thank an anonymous reviewer for this suggestion). This would lead to sequences such as ‘A<sub>1</sub> A<sub>2</sub> A<sub>3</sub> D B<sub>3</sub> B<sub>2</sub> B<sub>1</sub>’ for nested dependencies and sequences such as ‘A<sub>1</sub> A<sub>2</sub> A<sub>3</sub> D B<sub>1</sub> B<sub>2</sub> B<sub>3</sub>’ for cross-serial dependencies. According to the Chomsky Hierarchy adding a dummy variable would not affect predictions with respect to nested and cross-serial dependencies. Also the SPLT would not predict that adding a dummy variable would affect the memory load for nested dependencies in a different way than the memory load for cross-serial dependencies. In order to draw definite conclusions this empirical question would need to be addressed in a future study. If adding a dummy variable would lead to a higher performance for cross-serial compared to nested dependencies one would however have to incorporate findings by Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>], who found higher performance for cross-serial than for nested dependencies without having an intervening dummy variable.</p>
<p>In this context it is interesting to note that there is one difference between our study and that of Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] which could be made responsible for the different results, namely language experience [<xref rid="pone.0123059.ref030" ref-type="bibr">30</xref>]. The study by Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] investigated language learning in Dutch participants. Since Dutch, as Swiss German, contains grammatical constructions with cross-serial dependencies, speakers of Dutch are experienced in processing this type of dependencies [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]. The obtained processing advantage for cross-serial dependencies over nested dependencies in the study by Uddén and colleagues [<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>] could therefore be attributed to the fact that Dutch participants were investigated [<xref rid="pone.0123059.ref030" ref-type="bibr">30</xref>]. De Vries and colleagues [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>] compared the processing of nested and cross-serial dependencies in Dutch and German participants, as speakers of German, in contrast to Dutch speakers, should be more experienced in processing nested dependencies. In contrast to the predictions, their findings did not speak for a strong influence of language experience. Still, it could be the case that in the current study, language experience compensated the general processing advantage of cross-serial dependencies, because participants were native speakers of German (with the exception of one participant in Experiment 1). This hypothesis is strengthened by findings from Rohrmeier and colleagues [<xref rid="pone.0123059.ref031" ref-type="bibr">31</xref>]. In this study, language experience affected learning performance in an AGL experiment when acquiring dependencies of the context-free complexity level. Also Gervain and colleagues [<xref rid="pone.0123059.ref032" ref-type="bibr">32</xref>] showed that language experience affected performance of participants in an AGL experiment. To what degree language experience plays a role in language processing is an interesting question in the context of language learning and should be investigated further in future research.</p>
<p>In addition, it is of course possible that the differences with respect to learning performance appeared due to the different language material or the different control task in the learning phase. In our study, participants had to repeat aloud the last sequence in 25% of the trials in the learning phase to ensure that they were paying attention to the stimuli. Anecdotally, some participants reported after the end of the experiment that they found this task very difficult. Therefore, it could be the case that participants in the learning phase developed coping mechanisms in order to perform well in the control task of the learning phase. These coping mechanisms could have been beneficial with respect to extracting a rule in the test phase and thus might have contributed to the absence of performance differences between dependencies. Irrespective of this, the present study extends previous findings ([<xref rid="pone.0123059.ref021" ref-type="bibr">21</xref>], [<xref rid="pone.0123059.ref022" ref-type="bibr">22</xref>]) in that it shows that language learning can take place in a classical AGL setting within as little time as one hour and in a more natural language environment involving a larger stimulus set and the sequential presentation of spoken auditory stimuli.</p>
<p>Furthermore, we would like to note that a growing body of research suggests that a rule based system similar to the one discussed here for syntax also underlies phonology (for a review, see [<xref rid="pone.0123059.ref033" ref-type="bibr">33</xref>]). For studies employing the artificial grammar paradigm, in which an artificial language is being taught in the absence of semantic information, it would then be difficult to tell whether what is being investigated is syntax or phonology. In any case, considerations concerning complexity differences between the different types of dependencies investigated in the current study should hold independent of whether these dependencies are phonological or syntactic in nature. However, future studies are clearly needed that try to disentangle phonology from syntax in the artificial grammar paradigm. Investigations comparing the learning of the different rule based systems would certainly constitute an important step towards a better understanding of the cognitive processes involved in language processing.</p>
<p>Finally, in the present study, we only investigated two types of dependencies, and did so under restricted learning conditions. Future studies could take this study as a starting point for further exploring cognitive processing of complex dependencies in artificial languages and comparing it with predictions from the Chomsky Hierarchy and memory models such as the SPLT. The limit of learnability with respect to language complexity in the Chomsky Hierarchy would be particularly interesting. Thus, by assessing the learnability of dependencies located higher than cross-serial dependencies in the Chomsky Hierarchy, cognitive processing limits could be identified and could be compared with predictions by the SPLT. This would shed further light onto the compatibility of cognitive and formal complexity.</p>
</sec>
<sec id="sec021" sec-type="conclusions">
<title>Conclusion</title>
<p>Findings of the current study demonstrate the learnability of nested dependencies (context-free) and cross-serial dependencies (mildly context-sensitive) in a more natural AGL setting (spoken auditory stimuli, sequential presentation, larger stimulus set) after only one hour of language exposure. Differences between the dependency types did not become evident. Thus, the current finding suggests that formal complexity and cognitive complexity are not two sides of the same coin. Therefore, this study provides further empirical evidence for the conclusion recently drawn by Chesi and Moro [<xref rid="pone.0123059.ref026" ref-type="bibr">26</xref>] that the Chomsky Hierarchy does not reflect cognitive processes. Cognitive processing limits might become evident when investigating formally more complex dependencies.</p>
</sec>
</body>
<back>
<ack>
<p>We acknowledge support by Deutsche Forschungsgemeinschaft and Open Access Publishing Fund of University of Tübingen.</p>
</ack>
<ref-list>
<title>References</title>
<ref id="pone.0123059.ref001"><label>1</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Evans</surname> <given-names>N</given-names></name>, <name name-style="western"><surname>Levinson</surname> <given-names>SC</given-names></name>. <article-title>The myth of language universals: Language diversity and its importance for cognitive science</article-title>. <source>Behav Brain Sci</source>. <year>2009</year>;<volume>32</volume>:<fpage>429</fpage>–<lpage>48</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1017/S0140525X0999094X" xlink:type="simple">10.1017/S0140525X0999094X</ext-link></comment> <object-id pub-id-type="pmid">19857320</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref002"><label>2</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Chomsky</surname> <given-names>N</given-names></name>. <source>Syntactic structure</source>. <publisher-loc>The Hague</publisher-loc>: <publisher-name>Mouton</publisher-name>; <year>1957</year>.</mixed-citation></ref>
<ref id="pone.0123059.ref003"><label>3</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Chomsky</surname> <given-names>N</given-names></name>. <source>Aspects of the theory of syntax</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>MIT Press</publisher-name>; <year>1965</year>.</mixed-citation></ref>
<ref id="pone.0123059.ref004"><label>4</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Chomsky</surname> <given-names>N</given-names></name>. <article-title>Three models for the description of language</article-title>. <source>IRE Trans Inf Theory</source>. <year>1956</year>;<volume>2</volume>:<fpage>113</fpage>–<lpage>24</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref005"><label>5</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Jäger</surname> <given-names>G</given-names></name>, <name name-style="western"><surname>Rogers</surname> <given-names>J</given-names></name>. <article-title>Formal language theory: Refining the Chomsky hierarchy</article-title>. <source>Philos Trans R Soc Lond B Biol Sci</source>. <year>2012</year>;<volume>367</volume>:<fpage>1956</fpage>–<lpage>70</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1098/rstb.2012.0077" xlink:type="simple">10.1098/rstb.2012.0077</ext-link></comment> <object-id pub-id-type="pmid">22688632</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref006"><label>6</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Everett</surname> <given-names>DL</given-names></name>. <article-title>Cultural constraints on grammar and cognition in Pirahã: Another look at the design features of human language</article-title>. <source>Curr Anthropol</source>. <year>2005</year>;<volume>46</volume>:<fpage>621</fpage>–<lpage>46</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref007"><label>7</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Nevins</surname> <given-names>A</given-names></name>, <name name-style="western"><surname>Pesetsky</surname> <given-names>D</given-names></name>, <name name-style="western"><surname>Rodrigues</surname> <given-names>C</given-names></name>. <article-title>Pirahã exceptionality: A reassessment</article-title>. <source>Language</source>. <year>2009</year>;<volume>85</volume>:<fpage>355</fpage>–<lpage>404</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref008"><label>8</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Huybregts</surname> <given-names>R</given-names></name>. <chapter-title>The weak inadequacy of context-free phrase structure grammars</chapter-title>. In: <name name-style="western"><surname>de Haan</surname> <given-names>GJ</given-names></name>, <name name-style="western"><surname>Trommelen</surname> <given-names>M</given-names></name>, <name name-style="western"><surname>Zonneveld</surname> <given-names>W</given-names></name>, editors. <source>Van periferie naar kern</source>. <publisher-loc>Dordrecht</publisher-loc>: <publisher-name>Foris Publications</publisher-name>; <year>1984</year>. pp. <fpage>81</fpage>–<lpage>99</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref009"><label>9</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Joshi</surname> <given-names>AK</given-names></name>, <name name-style="western"><surname>Shanker</surname> <given-names>KV</given-names></name>, <name name-style="western"><surname>Weir</surname> <given-names>D</given-names></name>. <chapter-title>The convergence of mildly context-sensitive grammar formalisms</chapter-title>. In: <name name-style="western"><surname>Sells</surname> <given-names>P</given-names></name>, <name name-style="western"><surname>Shieber</surname> <given-names>S</given-names></name>, <name name-style="western"><surname>Wasow</surname> <given-names>T</given-names></name>, editors. <source>Foundational issues in natural language processing</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>MIT Press</publisher-name>; <year>1991</year>. pp. <fpage>31</fpage>–<lpage>81</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref010"><label>10</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Fitch</surname> <given-names>WT</given-names></name>, <name name-style="western"><surname>Hauser</surname> <given-names>MD</given-names></name>. <chapter-title>Computational constraints on syntactic processing in a nonhuman primate</chapter-title>. <source>Science</source>. <year>2004</year>;<volume>303</volume>:<fpage>377</fpage>–<lpage>80</lpage>. <object-id pub-id-type="pmid">14726592</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref011"><label>11</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Friederici</surname> <given-names>AD</given-names></name>, <name name-style="western"><surname>Bahlmann</surname> <given-names>J</given-names></name>, <name name-style="western"><surname>Heim</surname> <given-names>S</given-names></name>, <name name-style="western"><surname>Schubotz</surname> <given-names>RI</given-names></name>, <name name-style="western"><surname>Anwander</surname> <given-names>A</given-names></name>. <article-title>The brain differentiates human and non-human grammars: Functional localization and structural connectivity</article-title>. <source>Proc Natl Acad Sci U S A</source>. <year>2006</year>;<volume>103</volume>:<fpage>2458</fpage>–<lpage>63</lpage>. <object-id pub-id-type="pmid">16461904</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref012"><label>12</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Gasarch</surname> <given-names>W</given-names></name>. <chapter-title>Classifying problems into complexity classes</chapter-title>. In: <name name-style="western"><surname>Hurson</surname> <given-names>A</given-names></name>, <name name-style="western"><surname>Memon</surname> <given-names>A</given-names></name>, series editors. <source>Advances in computers</source>: <volume>Vol. 95</volume>. <publisher-loc>Waltham</publisher-loc>: <publisher-name>Academic Press</publisher-name>; <year>2014</year>. pp. <fpage>239</fpage>–<lpage>92</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref013"><label>13</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Alonso</surname> <given-names>MA</given-names></name>, <name name-style="western"><surname>Díaz</surname> <given-names>VJ</given-names></name>. <article-title>Variants of mixed parsing of TAG and TIG</article-title>. <source>Traitement Automatique des Langues (TAL)</source>. <year>2003</year>;<volume>44</volume>:<fpage>41</fpage>–<lpage>65</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref014"><label>14</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Joshi</surname> <given-names>AK</given-names></name>. <chapter-title>Tree Adjoining Grammars: How much context-sensitivity is necessary for characterizing structural descriptions?</chapter-title> In: <name name-style="western"><surname>Dowty</surname> <given-names>DR</given-names></name>, <name name-style="western"><surname>Karttunen</surname> <given-names>L</given-names></name>, <name name-style="western"><surname>Zwicky</surname> <given-names>AM</given-names></name>, editors. <source>Natural language parsing Psychological, computational and theoretical perspectives</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>; <volume>1985</volume>. <fpage>206</fpage>–<lpage>250</lpage>. Originally presented in a Workshop on Natural Language Parsing at Ohio State University, Columbus, Ohio, May 1983.</mixed-citation></ref>
<ref id="pone.0123059.ref015"><label>15</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Miller</surname> <given-names>GA</given-names></name>, <name name-style="western"><surname>McKean</surname> <given-names>KO</given-names></name>. <article-title>A chronometric study of some relations between sentences</article-title>. <source>Q J Exp Psychol</source>. <year>1964</year>;<volume>16</volume>:<fpage>297</fpage>–<lpage>308</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref016"><label>16</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Savin</surname> <given-names>HB</given-names></name>, <name name-style="western"><surname>Perchonock</surname> <given-names>E</given-names></name>. <article-title>Grammatical structure and the immediate recall of English sentences</article-title>. <source>J Verbal Learning Verbal Behav</source>. <year>1965</year>;<volume>4</volume>:<fpage>348</fpage>–<lpage>53</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref017"><label>17</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Reber</surname> <given-names>AS</given-names></name>. <article-title>Implicit learning of artificial grammars</article-title>. <source>J Verbal Learning Verbal Behav</source>. <year>1967</year>;<volume>6</volume>:<fpage>855</fpage>–<lpage>63</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref018"><label>18</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Perruchet</surname> <given-names>P</given-names></name>, <name name-style="western"><surname>Rey</surname> <given-names>A</given-names></name>. <article-title>Does the mastery of center-embedded linguistic structures distinguish humans from nonhuman primates?</article-title> <source>Psychon Bull Rev</source>. <year>2005</year>;<volume>12</volume>:<fpage>307</fpage>–<lpage>13</lpage>. <object-id pub-id-type="pmid">16082811</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref019"><label>19</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>de Vries</surname> <given-names>MH</given-names></name>, <name name-style="western"><surname>Monaghan</surname> <given-names>P</given-names></name>, <name name-style="western"><surname>Knecht</surname> <given-names>S</given-names></name>, <name name-style="western"><surname>Zwitserlood</surname> <given-names>P</given-names></name>. <article-title>Syntactic structure and artificial grammar learning: The learnability of embedded hierarchical structures</article-title>. <source>Cognition</source>. <year>2008</year>;<volume>107</volume>:<fpage>763</fpage>–<lpage>74</lpage>. <object-id pub-id-type="pmid">17963740</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref020"><label>20</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Hochmann</surname> <given-names>JR</given-names></name>, <name name-style="western"><surname>Azadpour</surname> <given-names>M</given-names></name>, <name name-style="western"><surname>Mehler</surname> <given-names>J</given-names></name>. <article-title>Do humans really learn A<sup>n</sup>B<sup>n</sup> artificial grammars from exemplars?</article-title> <source>Cogn Sci</source>. <year>2008</year>;<volume>32</volume>:<fpage>1021</fpage>–<lpage>36</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1080/03640210801897849" xlink:type="simple">10.1080/03640210801897849</ext-link></comment> <object-id pub-id-type="pmid">21585440</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref021"><label>21</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Uddén</surname> <given-names>J</given-names></name>, <name name-style="western"><surname>Ingvar</surname> <given-names>M</given-names></name>, <name name-style="western"><surname>Hagoort</surname> <given-names>P</given-names></name>, <name name-style="western"><surname>Petersson</surname> <given-names>KM</given-names></name>. <article-title>Implicit acquisition of grammars with crossed and nested non-adjacent dependencies: Investigating the push-down stack model</article-title>. <source>Cogn Sci</source>. <year>2012</year>;<volume>36</volume>:<fpage>1078</fpage>–<lpage>101</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1111/j.1551-6709.2012.01235.x" xlink:type="simple">10.1111/j.1551-6709.2012.01235.x</ext-link></comment> <object-id pub-id-type="pmid">22452530</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref022"><label>22</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>de Vries</surname> <given-names>MH</given-names></name>, <name name-style="western"><surname>Petersson</surname> <given-names>KM</given-names></name>, <name name-style="western"><surname>Geukes</surname> <given-names>S</given-names></name>, <name name-style="western"><surname>Zwitserlood</surname> <given-names>P</given-names></name>, <name name-style="western"><surname>Christiansen</surname> <given-names>MH</given-names></name>. <article-title>Processing multiple non-adjacent dependencies: Evidence from sequence learning</article-title>. <source>Philos Trans R Soc Lond B Biol Sci</source>. <year>2012</year>;<volume>367</volume>:<fpage>2065</fpage>–<lpage>76</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1098/rstb.2011.0414" xlink:type="simple">10.1098/rstb.2011.0414</ext-link></comment> <object-id pub-id-type="pmid">22688641</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref023"><label>23</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Bahlmann</surname> <given-names>J</given-names></name>, <name name-style="western"><surname>Gunter</surname> <given-names>TC</given-names></name>, <name name-style="western"><surname>Friederici</surname> <given-names>AD</given-names></name>. <article-title>Hierarchical and linear sequence processing: An electrophysiological exploration of two different grammar types</article-title>. <source>J Cogn Neurosci</source>. <year>2006</year>;<volume>18</volume>:<fpage>1829</fpage>–<lpage>42</lpage>. <object-id pub-id-type="pmid">17069474</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref024"><label>24</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Bahlmann</surname> <given-names>J</given-names></name>, <name name-style="western"><surname>Schubotz</surname> <given-names>RI</given-names></name>, <name name-style="western"><surname>Friederici</surname> <given-names>AD</given-names></name>. <article-title>Hierarchical artificial grammar processing engages Broca's area</article-title>. <source>Neuroimage</source>. <year>2008</year>;<volume>42</volume>:<fpage>525</fpage>–<lpage>34</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1016/j.neuroimage.2008.04.249" xlink:type="simple">10.1016/j.neuroimage.2008.04.249</ext-link></comment> <object-id pub-id-type="pmid">18554927</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref025"><label>25</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Friederici</surname> <given-names>AD</given-names></name>. <article-title>Processing local transitions versus long-distance syntactic hierarchies</article-title>. <source>Trends Cogn Sci</source>. <year>2004</year>;<volume>8</volume>:<fpage>245</fpage>–<lpage>7</lpage>. <object-id pub-id-type="pmid">15165545</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref026"><label>26</label><mixed-citation publication-type="book" xlink:type="simple"><name name-style="western"><surname>Chesi</surname> <given-names>C</given-names></name>, <name name-style="western"><surname>Moro</surname> <given-names>A</given-names></name>. <chapter-title>Computational complexity in the brain</chapter-title>. In: <name name-style="western"><surname>Newmeyer</surname> <given-names>FJ</given-names></name>, <name name-style="western"><surname>Preston</surname> <given-names>LB</given-names></name>, editors. <source>Measuring grammatical complexity</source>, <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>; <year>2014</year>. pp. <fpage>264</fpage>–<lpage>280</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref027"><label>27</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Gibson</surname> <given-names>E</given-names></name>. <article-title>Linguistic complexity: Locality of syntactic dependencies</article-title>. <source>Cognition</source>. <year>1998</year>;<volume>68</volume>:<fpage>1</fpage>–<lpage>76</lpage>. <object-id pub-id-type="pmid">9775516</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref028"><label>28</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Bach</surname> <given-names>E</given-names></name>, <name name-style="western"><surname>Brown</surname> <given-names>C</given-names></name>, <name name-style="western"><surname>Marslen-Wilson</surname> <given-names>W</given-names></name>. <article-title>Crossed and nested dependencies in German and Dutch: A psycholinguistic study</article-title>. <source>Lang Cogn Process</source>. <year>1986</year>;<volume>1</volume>:<fpage>249</fpage>–<lpage>62</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref029"><label>29</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Hautus</surname> <given-names>MJ</given-names></name>. <article-title>Corrections for extreme proportions and their biasing effects on estimated values of d′</article-title>. <source>Behav Res Methods Instrum Comput</source>. <year>1995</year>;<volume>27</volume>:<fpage>46</fpage>–<lpage>51</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref030"><label>30</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>de Vries</surname> <given-names>MH</given-names></name>, <name name-style="western"><surname>Christiansen</surname> <given-names>MH</given-names></name>, <name name-style="western"><surname>Petersson</surname> <given-names>KM</given-names></name>. <article-title>Learning recursion: Multiple nested and crossed dependencies</article-title>. <source>Biolinguistics</source>. <year>2011</year>;<volume>5</volume>:<fpage>10</fpage>–<lpage>35</lpage>.</mixed-citation></ref>
<ref id="pone.0123059.ref031"><label>31</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Rohrmeier</surname> <given-names>M</given-names></name>, <name name-style="western"><surname>Fu</surname> <given-names>Q</given-names></name>, <name name-style="western"><surname>Dienes</surname> <given-names>Z</given-names></name>. <article-title>Implicit learning of recursive context-free grammars</article-title>. <source>PLoS One</source>. <year>2012</year>;<volume>7</volume>:<fpage>e45885</fpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1371/journal.pone.0045885" xlink:type="simple">10.1371/journal.pone.0045885</ext-link></comment> <object-id pub-id-type="pmid">23094021</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref032"><label>32</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Gervain</surname> <given-names>J</given-names></name>, <name name-style="western"><surname>Sebastián-Gallés</surname> <given-names>N</given-names></name>, <name name-style="western"><surname>Díaz</surname> <given-names>B</given-names></name>, <name name-style="western"><surname>Laka</surname> <given-names>I</given-names></name>, <name name-style="western"><surname>Mazuka</surname> <given-names>R</given-names></name>, <name name-style="western"><surname>Yamane</surname> <given-names>N</given-names></name>, <etal>et al</etal>. <article-title>Word frequency cues word order in adults: Cross-linguistic evidence</article-title>. <source>Front Psychol</source>. <year>2013</year>;<volume>4</volume>:<fpage>689</fpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.3389/fpsyg.2013.00689" xlink:type="simple">10.3389/fpsyg.2013.00689</ext-link></comment> <object-id pub-id-type="pmid">24106483</object-id></mixed-citation></ref>
<ref id="pone.0123059.ref033"><label>33</label><mixed-citation publication-type="journal" xlink:type="simple"><name name-style="western"><surname>Berent</surname> <given-names>I</given-names></name>. <article-title>The phonological mind</article-title>. <source>Trends Cogn Sci</source>. <year>2013</year>;<volume>17</volume>:<fpage>319</fpage>–<lpage>27</lpage>. <comment>doi: <ext-link ext-link-type="uri" xlink:href="http://dx.doi.org/10.1016/j.tics.2013.05.004" xlink:type="simple">10.1016/j.tics.2013.05.004</ext-link></comment> <object-id pub-id-type="pmid">23768723</object-id></mixed-citation></ref>
</ref-list>
</back>
</article>