<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2023.1246710</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>What corpus data reveal about the Position of Antecedent Strategy: anaphora resolution in Spanish monolinguals and L1 English-L2 Spanish bilinguals</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name><surname>Lozano</surname> <given-names>Crist&#x00F3;bal</given-names></name>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/1832325/overview"/>
</contrib>
<contrib contrib-type="author">
<name><surname>Quesada</surname> <given-names>Teresa</given-names></name>
<uri xlink:href="https://loop.frontiersin.org/people/2301116/overview"/>
</contrib>
</contrib-group>
<aff><institution>Universidad de Granada</institution>, <addr-line>Granada</addr-line>, <country>Spain</country></aff>
<author-notes>
<fn fn-type="edited-by" id="fn0014">
<p>Edited by: Tania L. Leal, University of Arizona, United States</p>
</fn>
<fn fn-type="edited-by" id="fn0015">
<p>Reviewed by: Tiffany Judy, Wake Forest University, United States; Tihana Kras, University of Rijeka, Croatia</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: Crist&#x00F3;bal Lozano, <email>cristoballozano@ugr.es</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>09</day>
<month>11</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1246710</elocation-id>
<history>
<date date-type="received">
<day>24</day>
<month>06</month>
<year>2023</year>
</date>
<date date-type="accepted">
<day>18</day>
<month>10</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2023 Lozano and Quesada.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Lozano and Quesada</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>This study investigates the acquisition of anaphora resolution (AR) in Spanish as a second language (L2). According to the Position of Antecedent Strategy (PAS), in native Spanish null pronominal subjects are biased toward subject antecedents, whereas overt pronominal subjects show a &#x201C;flexible&#x201D; bias (typically toward non-subject but also toward subject antecedents). The PAS has been extensively investigated in experimental studies, though little is known about real production. We show how naturalistic production (corpus methods) can uncover crucial factors in the PAS that have not been explored in the experimental literature. We analyzed written samples from the CEDEL2 corpus: L1 English-L2 Spanish adult late-bilingual learners (intermediate, lower-advanced and upper-advanced proficiency levels) and a control group of adult Spanish monolinguals (<italic>N</italic>&#x2009;=&#x2009;75 texts). Anaphors were manually annotated via a fine-grained, linguistically-motivated tagset in UAM Corpus Tool. Against traditional assumptions, our results reveal that (i) the PAS is not a privileged mechanism for resolving anaphora; (ii) it is more complex than assumed (in terms of the division of labor of anaphoric forms, their antecedents and the syntactic configuration in which they appear); (iii) the much-debated &#x201C;flexible&#x201D; bias of overt pronouns is apparent since they are hardly produced and are replaced by repeated NPs, which show a clear non-subject antecedent bias; (iv) at the syntax-discourse interface, the PAS is constrained by information structure in more complex ways than assumed: null pronouns mark topic continuity, whereas overtly realized referential expressions (overt REs: overt pronouns and NPs) mark topic shift. Learners show more difficulties with topic continuity (where they redundantly use overt pronouns) than with topic shift (where they normally disambiguate by using overtly realized REs), thus being more redundant than ambiguous, in line with the Pragmatic Principles Violation Hypothesis (PPVH) (Lozano, 2016). We finally argue that the insights from corpora should be implemented into experiments. The triangulation of corpus and experimental methods in bilingualism ultimately provides a clearer understanding of the phenomenon under investigation.</p>
</abstract>
<kwd-group>
<kwd>Spanish second language acquisition</kwd>
<kwd>anaphora resolution</kwd>
<kwd>position of antecedent strategy</kwd>
<kwd>learner corpora</kwd>
<kwd>pronominal subjects</kwd>
<kwd>CEDEL2 corpus</kwd>
</kwd-group>
<counts>
<fig-count count="12"/>
<table-count count="2"/>
<equation-count count="0"/>
<ref-count count="45"/>
<page-count count="21"/>
<word-count count="13160"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Psychology of Language</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="sec1">
<label>1.</label>
<title>Introduction: anaphora resolution and the Position of Antecedent Strategy</title>
<p>Anaphora Resolution (AR) is a frequent and pervasive (though deceptively simple) mechanism found in all natural languages. Its acquisition represents a challenge for different types of bilinguals, including late sequential bilinguals like adult second language (L2) learners (<xref ref-type="bibr" rid="ref28">Lozano, 2021a</xref>).</p>
<p>Anaphors like pronominal subjects refer to their antecedents in discourse. The ambiguous scenario in English (1) requires the resolution of the anaphor: <italic>she</italic> can refer to either antecedent (subject <italic>Carmen</italic> or object <italic>Paola</italic>). Null-subject languages like Spanish are anaphorically more complex since both null (&#x00D8;) and overt (<italic>ella</italic> &#x201C;she&#x201D;) pronouns can alternate in subject syntactic position, (2), and can refer to either antecedent. Despite this apparent ambiguity, our mental syntactic parser/processor has certain strategies to automatically resolve the anaphor.</p>
<list list-type="simple">
<list-item>
<p>(1) Carmen<sub>i</sub> greeted Paola<sub>j</sub> while <bold>she</bold><sub>
<bold>i/j</bold>
</sub> was opening the door.</p>
</list-item>
<list-item>
<p>(2) Carmen<sub>i</sub> salud&#x00F3; a Paola<sub>j</sub> mientras <inline-formula>
<mml:math id="M1">
<mml:mrow>
<mml:mfenced close="}" open="{">
<mml:mrow>
<mml:mtable equalrows="true" equalcolumns="true">
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mstyle mathvariant="bold">
<mml:mi>&#x00D8;</mml:mi>
</mml:mstyle>
<mml:mrow>
<mml:mstyle mathvariant="bold">
<mml:mi>i</mml:mi>
</mml:mstyle>
<mml:mi mathvariant="normal">/</mml:mi>
<mml:mstyle mathvariant="bold">
<mml:mi>j</mml:mi>
</mml:mstyle>
</mml:mrow>
</mml:msub>
<mml:mspace width="thickmathspace"/>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mstyle mathvariant="bold">
<mml:mi>e</mml:mi>
<mml:mi>l</mml:mi>
<mml:mi>l</mml:mi>
</mml:mstyle>
<mml:msub>
<mml:mstyle mathvariant="bold">
<mml:mi>a</mml:mi>
</mml:mstyle>
<mml:mrow>
<mml:mstyle mathvariant="bold">
<mml:mi>i</mml:mi>
</mml:mstyle>
<mml:mi mathvariant="normal">/</mml:mi>
<mml:mstyle mathvariant="bold">
<mml:mi>j</mml:mi>
</mml:mstyle>
</mml:mrow>
</mml:msub>
<mml:mspace width="thickmathspace"/>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:math>
</inline-formula> abr&#x00ED;a la puerta.</p>
</list-item>
</list>
<p>The Position of Antecedent Strategy (PAS),<xref ref-type="fn" rid="fn0001"><sup>1</sup></xref> originally formulated by <xref ref-type="bibr" rid="ref5">Carminati (2002)</xref> for native Italian, resolves such ambiguity in intrasentential AR (subordinate-main clausal order). Carminati&#x2019;s results from an offline sentence-interpretation task confirmed this trend: When asked about the interpretation of the second clause (e.g., <italic>Who was in the United States?</italic>), Italian monolinguals chose a subject antecedent (<italic>Marta</italic> 80.72%) with null pronouns in (3), but a non-subject antecedent (<italic>Piera</italic> 83.33%) with overt pronouns. Results from an online self-paced reading task (SPRT) confirmed this: null pronominals (<italic>&#x00D8;</italic>) take significantly shorter when referring to preverbal subjects (1,844&#x2009;ms) than to postverbal objects (2,352&#x2009;ms), whereas overt pronouns (<italic>lei</italic> &#x201C;she&#x201D;) take less time to non-subject (2,236&#x2009;ms) than to subject (2,266&#x2009;ms) antecedents.</p>
<list list-type="simple">
<list-item>
<p>(3) Marta scriveva frequentemente a Piera quando <inline-formula>
<mml:math id="M2">
<mml:mrow>
<mml:mfenced close="}" open="{">
<mml:mrow>
<mml:mtable equalrows="true" equalcolumns="true">
<mml:mtr>
<mml:mtd>
<mml:mstyle mathvariant="bold">
<mml:mi>&#x00D8;</mml:mi>
</mml:mstyle>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mstyle mathvariant="bold">
<mml:mi>l</mml:mi>
</mml:mstyle>
<mml:msub>
<mml:mstyle mathvariant="bold">
<mml:mi>e</mml:mi>
</mml:mstyle>
<mml:mstyle mathvariant="bold">
<mml:mi>i</mml:mi>
</mml:mstyle>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
</mml:mfenced>
</mml:mrow>
</mml:math>
</inline-formula> era negli Stati Uniti.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Marta wrote frequently to Piera when &#x00D8;/she was in the United States.&#x201D;</p>
</disp-quote>
<p>The PAS is a syntactic/configurational parsing strategy with a clear division of labor: null pronouns are biased toward a preverbal subject antecedent whereas overt pronouns are biased toward a postverbal non-subject antecedent. Importantly, the PAS is also a syntax-discourse interface phenomenon due to the information status of the anaphor: null pronouns encode a continuation of the preceding subject (topic continuity), whereas overt pronouns mark a topic shift. This holds true in other null-subject languages like Spanish (<xref ref-type="bibr" rid="ref25">Lozano, 2009</xref>, <xref ref-type="bibr" rid="ref26">2016</xref>, <xref ref-type="bibr" rid="ref28">2021a</xref>; <xref ref-type="bibr" rid="ref34">Mart&#x00ED;n-Villena and Lozano, 2020</xref>), Moroccan Arabic (<xref ref-type="bibr" rid="ref2">Bel and Garc&#x00ED;a-Alcaraz, 2015</xref>), Greek (<xref ref-type="bibr" rid="ref40">Prentza and Tsimpli, 2013</xref>; <xref ref-type="bibr" rid="ref38">Papadopoulou et al., 2015</xref>), Croatian (<xref ref-type="bibr" rid="ref23">Kra&#x0161;, 2008a</xref>,<xref ref-type="bibr" rid="ref24">b</xref>), and Romanian (<xref ref-type="bibr" rid="ref17">Geber, 2006</xref>), among other languages.</p>
<p>The PAS had been extensively investigated in diverse bilingual populations (adult and child L2 learners, heritage speakers, attriters) in different L1-L2 combinations, which has led to the proposal of key theories like <xref ref-type="bibr" rid="ref44">Sorace&#x2019;s (2011)</xref> Interface Hypothesis (IH), which predicts bilinguals to show limitations when simultaneously integrating syntactic and discursive information. Follow-up proposals, like <xref ref-type="bibr" rid="ref26">Lozano&#x2019;s (2016)</xref> Pragmatic Principles Violation Hypothesis (PPVH), locate the source limitations at a more pragmatic level (topic continuity vs. shift), as a result of the violation of pragmatic principles like Economy vs. Clarity.</p>
<p>Crucially, much of our understanding of AR in general and PAS in particular comes from experimental studies that (i) often report contradictory results, so it is still unclear how the PAS operates in native (and L2) Spanish, and (ii) repeatedly investigate similar anaphoric configuration (i.e., PAS). We argue that highly-contextualized, discourse-rich corpus production data can uncover many factors that have gone undetected in prior experimental studies and solves some of the unresolved PAS questions. Additionally, our developmental corpus data will also allow us to know how the PAS is acquired across proficiency in L1 English-L2 Spanish and whether very advanced learners can eventually acquire the pragmatic subtleties of PAS.</p>
<p>Carminati&#x2019;s PAS was originally formulated for language processing (comprehension) and our aim is to put it to the test in language production (corpus data). In the psycholinguistic literature, it has long been acknowledged that &#x201C;grammatical processing (or &#x201C;parsing&#x201D;) &#x2026; refers to the construction of structural representations for sentences, phrases and morphologically complex words in real-time language comprehension and production&#x201D; (<xref ref-type="bibr" rid="ref7">Clahsen and Felser, 2006</xref>, p. 564) and that &#x201C;there may be a closer link between comprehension and production, in particular between parsing and syntactic encoding during production.&#x201D; (<xref ref-type="bibr" rid="ref39">Pickering and van Gompel, 2006</xref>, p. 487). In this line, <xref ref-type="bibr" rid="ref31">Mac Donald (2013)</xref> empirically shows that &#x201C;language production processes can provide insight into how language comprehension works&#x201D; (p. 1) and concludes that &#x201C;the availability of extensive language corpora in many languages permits comprehension researchers to examine the relationship between production patterns (in the corpus) and comprehension behavior&#x201D; (p. 13). Additionally, it is widely acknowledged in the (bilingual) psycholinguistic literature (e.g., <xref ref-type="bibr" rid="ref11">Fern&#x00E1;ndez and Smith Cairns, 2011</xref>) that, during processing (parsing), two major processes take place: (i) structuring the incoming input into categories, and (ii) establishing appropriate dependency relations between such categories, which is particularly relevant when there is potential ambiguity (as is the case in PAS scenarios). AR in general and the PAS in particular are classic examples of dependency. Dependencies need to be established not only in comprehension (listener/reader&#x2019;s perspective), but also in production since the speaker/writer needs to make sure that the anaphoric dependency s/he is producing is configurationally well established and structured (as is the case of PAS scenarios) to ensure that the listener/reader can interpret such dependency and therefore resolve the anaphor. Therefore, the use of production methods (corpora) can shed light on the PAS, as we do in this paper.</p>
<p>We next review the acquisition and processing of PAS in native and L2 Spanish based on experimental and corpus studies (section 1.1). In section 1.2 we present the research questions. The corpus methodology is discussed in section 2. Section 3 presents the results for each research question followed by a discussion, and section 4 concludes with a general discussion/conclusion and future avenues of investigation.</p>
<sec id="sec2">
<label>1.1.</label>
<title>The PAS in native and L2 Spanish</title>
<p>Overall, previous experimental native Spanish PAS findings show no clear division of labor as in native Italian: null pronouns select subject antecedents, but overt pronouns are &#x201C;flexible&#x201D; (non-subject and subject antecedents). Each experimental study is unique in terms of, e.g., the type of method/stimuli/design, which could explain the different results across studies. Consequently, we present a thorough review of each study to detect possible limitations that will be later implemented in our corpus study. Note that we review both offline and online PAS studies in adult Spanish monolinguals and adult L2 learners, thereby excluding other populations (see <xref ref-type="table" rid="tab1">Tables 1</xref>, <xref ref-type="table" rid="tab2">2</xref> in the <xref ref-type="supplementary-material" rid="SM1">online</xref> <xref ref-type="supplementary-material" rid="SM1">Supplementary material</xref> for additional details).<xref ref-type="fn" rid="fn0002"><sup>2</sup></xref> Finally, no single corpus study has targeted PAS structures, so we review some corpus evidence on AR in general as their findings may shed light on PAS.</p>
<sec id="sec3">
<label>1.1.1.</label>
<title>Offline experimental evidence</title>
<p><xref ref-type="bibr" rid="ref1">Alonso-Ovalle et al. (2002)</xref> administered a sentence interpretation task with intersentential PAS (4) to adult Peninsular Spanish monolinguals. Results from the comprehension question (Who is angry?) show a clear subject bias (<italic>Juan</italic> 73.2%) for null pronouns but a &#x201C;flexible&#x201D; behavior for overt pronouns (50.2% subject antecedent <italic>Juan</italic>, 49.8% non-subject antecedent <italic>Pedro</italic>), contra <xref ref-type="bibr" rid="ref5">Carminati&#x2019;s (2002)</xref> original PAS formulation.</p>
<list list-type="simple">
<list-item>
<p>(4)Juan peg&#x00F3; a Pedro. <inline-formula>
<mml:math id="M3">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi>&#x00E9;</mml:mi>
<mml:mi mathvariant="normal">l</mml:mi>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> est&#x00E1; enfadado.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Juan hit Pedro. (He) is angry.&#x201D;</p>
</disp-quote>
<p>Adult Peninsular Spanish monolinguals (with knowledge of Catalan) were tested in an acceptability judgment continuation task, where the plausibility of the continuation sentence (<italic>in italics</italic>) was judged on a four-point scale (<xref ref-type="bibr" rid="ref3">Bel et al., 2016a</xref>). Monolinguals judged main-subordinate clause order (5) vs. subordinate-main clause order (e.g., <italic>Mientras Javier abandonaba a Pedro, se emborrach&#x00F3;. Pedro se emborrach&#x00F3;</italic>).</p>
<list list-type="simple">
<list-item>
<p>(5) Javier abandon&#x00F3; a Pedro miembras se emborrachaba. <italic>Pedro se emborrachaba.</italic></p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Javier abandoned Pedro while (he) was getting drunk. Pedro was getting drunk.&#x201D;</p>
</disp-quote>
<p>When both clausal orders are analyzed together, null pronouns refer more to the subject (mean: 3.1) than the object (2.6), but overt pronouns refer to the object (3.2) more than to the subject (2.3). The same holds for <italic>subordinate-main</italic> order (null: 3.55 subject, 2.25 object; overt: 3.25 object, 2.45 subject). This confirms Carminatti&#x2019;s PAS. In <italic>main-subordinate</italic> order, results for the null pronoun were unexpected (null: 2.71 subject, 3.03 object; overt: object 3.01, subject 2.18). These unexpected monolingual finding led us to incorporate clausal order as a variable in our corpus-based study. The results for monolinguals were similar in <xref ref-type="bibr" rid="ref2">Bel and Garc&#x00ED;a-Alcaraz (2015)</xref>, who also included intermediate adult L1 Arabic-L2 Spanish learners in Morocco, both Moroccan Arabic and Spanish being null-subject languages with similar PAS behavior. Learners observed the PAS timidly in both clausal orders: (i) in <italic>main-subordinate</italic>, the null pronouns selected subjects (2.74 in main-subordinate, 2.64 in subordinate-main) slightly more than objects (2.54 and 2.34 respectively), but overt pronouns chose objects (2.81 and 2.63) more than subjects (2.16 and 2.40). In short, learners obey the PAS timidly, whereas Spanish(/Catalan) monolinguals do as well except for the main-subordinate condition, where null pronouns show the opposite behavior.</p>
<p>Jegerski and colleagues conducted a couple of PAS studies. First (<xref ref-type="bibr" rid="ref20">Jegerski et al., 2011</xref>), they tested L1 English-L2 Spanish adult learners (intermediate, advanced) and adult Spanish monolinguals (from Spain and Latin America) in an ambiguous PAS sentence-interpretation task with null and overt pronouns (6).<xref ref-type="fn" rid="fn0003"><sup>3</sup></xref></p>
<list list-type="simple">
<list-item>
<p>(6)Marta le escrib&#x00ED;a frecuentemente a Lorena cuando <inline-formula>
<mml:math id="M4">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi mathvariant="normal">ella</mml:mi>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> estaba en los Estados Unidos.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Marta wrote frequently to Lorena when (she) was in the United States&#x201D;.</p>
</disp-quote>
<p>When asked about the anaphoric interpretation, monolinguals preferred to link null pronouns with subject antecedents (75%), as predicted by the PAS, but overt pronouns show again a &#x201C;flexible&#x201D; behavior (53% subject antecedents, 47% object antecedents). Advanced learners show a native-like tendency: null-subject 69%, and &#x201C;flexible&#x201D; overt pronoun behavior (56% subject antecedent, 44% object antecedent). Intermediates show a timid subject bias irrespective of the pronoun type (null-subject 66%, overt-subject 60%). In their second study, <xref ref-type="bibr" rid="ref22">Keating et al. (2011)</xref> employed the same methodology and the same profiles of participants. Once again, Spanish monolinguals significantly preferred a null pronoun (74%) to an overt pronoun (54%) to refer to the subject. By contrast, the difference was not significant in advanced learners (60.15% null vs. 54.21% overt). Results from both studies indicate that overt pronouns show a &#x201C;flexible&#x201D; behavior by referring around 50% of the time to the subject and 50% to the object, both in native and L2 Spanish, a fact to which we will return in our study.</p>
<p>In a picture-verification task, <xref ref-type="bibr" rid="ref8">Clements and Dom&#x00ED;nguez (2017)</xref> tested the PAS in adult monolinguals (mainly from Spain, some from Mexico) and advanced L1 English-L2 Spanish learners from the United Kingdom, who were presented with two pictures and a PAS sentence with(out) an overt pronoun, as in (6). They had to decide whether the given sentence corresponded to one or the other picture (or both). Monolinguals preferred to link a null pronoun with a subject (77%) more than an object (12%) antecedent, whereas overt pronouns showed the opposite pattern (54% object, 27% subject), which supports Carminati&#x2019;s original PAS formulation, though note once again that the intuitions for overt pronouns are not as strong as those for null pronouns, a fact to which we will return in this paper. Unlike previous findings above, advanced learners observed the PAS in a native-like manner (null: subject 68%, object 21%; overt: object 63%, subject 23%).</p>
<p><xref ref-type="bibr" rid="ref6">Chamorro et al. (2016)</xref> asked adult monolinguals from Spain to rate null/overt pronoun PAS under four conditions: two forced antecedent-subject biases (singular subject, plural object (7a)), and two forced object-antecedent biases (plural subject, singular object (7b)). Monolinguals non-significantly rated the null pronoun to equally refer to the subject (3.72) and the object (3.61) antecedent, showing no clear subject bias of null pronouns, which runs against all the findings reviewed above. The overt pronoun significantly referred to the object (3.60) more than the subject (3.26) antecedent (though note the 3.26 vs. 3.60 ratings are not different enough given the 1&#x2013;5 Likert rating scale).</p>
<list list-type="simple">
<list-item>
<p>(7) a. La madre<sub>i</sub> salud&#x00F3; a las chicas<sub>j</sub> cuando <inline-formula>
<mml:math id="M5">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">ella</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> cruzaba una calle con mucho tr&#x00E1;fico.</p>
</list-item>
<list-item>
<p>b. Las madres<sub>i</sub> saludaron a la chica<sub>j</sub> cuando <inline-formula>
<mml:math id="M6">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">ella</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> cruzaba una calle con mucho tr&#x00E1;fico.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;The mother (s) greeted the girl (s) when (she) was crossing a street with lots of traffic&#x201D;.</p>
</disp-quote>
<p>In a picture selection task, <xref ref-type="bibr" rid="ref33">Mart&#x00ED;n-Villena (2023)</xref> tested conjunction type (when vs. while) in Peninsular Spanish monolinguals in sentences like (6). Subject-antecedent preferences with conjunction <italic>cuando</italic> &#x201C;when&#x201D; were higher for null (67%) than overt (23%) pronouns as well as with <italic>mientras</italic> &#x201C;while&#x201D; (null: 80%; overt: 30%). This confirms PAS preferences for subject antecedents but shows that null-subject bias was somewhat stronger with mientras &#x201C;while&#x201D; than with cuando &#x201C;when&#x201D;.</p>
</sec>
<sec id="sec4">
<label>1.1.2.</label>
<title>Online experimental evidence</title>
<p>All online experiments to date have used SPRT, which measure reading times (RTs) in milliseconds (ms). <xref ref-type="bibr" rid="ref12">Filiaci (2010)</xref> was the first online study to test PAS in Peninsular Spanish monolinguals. In intrasentential subordinate-main clauses, (8), the semantics of the main clause forced the anaphor toward the subject (8a) or the object (8b) antecedent. RTs of the main clause with a null pronoun were significantly faster when biasing toward the subject (1,998&#x2009;ms) than the object (2,319&#x2009;ms) antecedent, as predicted by Carminati&#x2019;s PAS, but with an overt pronoun, RTs were faster when biasing toward the object (2,389&#x2009;ms) than the subject (2,507&#x2009;ms) (but differences were non-significant, which reflects again the &#x201C;flexible&#x201D; behavior of Spanish overt pronouns). These results were later published (<xref ref-type="bibr" rid="ref13">Filiaci et al., 2014</xref>) as experiment 1. Experiment 2 stimuli were the same as in experiment 1 but RT analyses were conducted at different phrasal regions (separated by slashes &#x201C;/&#x201D; in (9)). Overall, results replicated those found in experiment 1, thus confirming the &#x201C;flexibility&#x201D; of overt pronouns in Spanish when compared to Italian.</p>
<list list-type="simple">
<list-item>
<p>(8) a. Cuando Ana<sub>i</sub> visit&#x00F3; a Mar&#x00ED;a<sub>j</sub> en en el hospital, <inline-formula>
<mml:math id="M7">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">ella</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> le llev&#x00F3; un ramo de rosas.</p>
</list-item>
<list-item>
<p>b. Cuando Ana<sub>i</sub> visit&#x00F3; a Mar&#x00ED;a<sub>j</sub> en en el hospital, <inline-formula>
<mml:math id="M8">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">ella</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> ya estaba fuera de peligro.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;When Ana visited Mary in the hospital, (she) {brought her a bunch of roses | was already out of danger.}&#x201D;</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>(9) Cuando / Ana / visit&#x00F3; / a Mar&#x00ED;a / en en el hospital, / <inline-formula>
<mml:math id="M9">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">ella</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> / le llev&#x00F3; / un ramo / de rosas.</p>
</list-item>
</list>
<p><xref ref-type="bibr" rid="ref18">Gelormini-Lezama and Almor (2011)</xref> tested intersentential PAS with adult Argentinian Spanish monolinguals. Sentences also included repeated names (RNs) (e.g., <italic>Juan</italic> &#x201C;John&#x201D;), (10). The object clitic (<italic>la</italic> &#x201C;her&#x201D;) forces the null pronoun toward a subject (10a) or object (10b) antecedent reading. With forced subject antecedents, RTs for null-pronoun sentences (1,812&#x2009;ms) were faster than overt-pronoun sentences (2264), but the opposite was true when with forced object antecedents (null 2,412, overt 2,157). This clearly confirms Carminati&#x2019;s PAS prediction. Interestingly, RNs were read equally fast irrespective of their antecedent (2080 subject, 2055 object) and their RTs did not significantly differ from sentences containing overt pronouns but did significantly differ from sentences containing null pronouns (subject: null &#x003C; RN; object: null &#x003E; RN), which suggests that NPs may play a role in object-antecedent selection in AR in native Spanish, a fact to which we will return in our corpus analysis.</p>
<list list-type="simple">
<list-item>
<p>(10) a. Juan<sub>i</sub> se encontr&#x00F3; con Mar&#x00ED;a<sub>j</sub>. <inline-formula>
<mml:math id="M10">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi>&#x00C9;</mml:mi>
<mml:msub>
<mml:mi mathvariant="normal">l</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">Juan</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> la<sub>j</sub> vio triste.</p>
</list-item>
<list-item>
<p>b. Mar&#x00ED;a<sub>i</sub> se encontr&#x00F3; con Juan<sub>j</sub>. <inline-formula>
<mml:math id="M11">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi>&#x00C9;</mml:mi>
<mml:msub>
<mml:mi mathvariant="normal">l</mml:mi>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mrow>
<mml:mi mathvariant="normal">Juan</mml:mi>
</mml:mrow>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> la<sub>i</sub> vio triste.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;{John found Mary | Mary found John}. &#x00D8;/He/John found her sad.&#x201D;</p>
</disp-quote>
<p>Another study (<xref ref-type="bibr" rid="ref4">Bel et al., 2016b</xref>) tested adult Peninsular Spanish monolinguals in intrasentential (main-subordinate order) PAS, (11), presented in a word-by-word, non-cumulative fashion. The ambiguous anaphor is resolved postverbally via world knowledge: <italic>violin</italic> forces a subject antecedent (musician), whereas <italic>casco</italic> &#x201C;helmet&#x201D; forces an object antecedent (firefighter).</p>
<list list-type="simple">
<list-item>
<p>(11) El m&#x00FA;sico<sub>i</sub> saluda al bombero<sub>j</sub> mientras <inline-formula>
<mml:math id="M12">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi>&#x00E9;</mml:mi>
<mml:msub>
<mml:mi mathvariant="normal">l</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> lleva <inline-formula>
<mml:math id="M13">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi mathvariant="normal">un</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mi mathvariant="normal">viol</mml:mi>
<mml:mi>&#x00ED;</mml:mi>
<mml:mi mathvariant="normal">n</mml:mi>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi mathvariant="normal">un</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mi mathvariant="normal">casco</mml:mi>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> en la mochila.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;The musician greets the fireman while (he) carries {a violin | a helmet} in his backpack.&#x201D;</p>
</disp-quote>
<p>Null pronouns were read significantly faster with a subject-antecedent bias (798&#x2009;ms) than an object-antecedent bias (887&#x2009;ms) at the NP object region (<italic>un violin/un casco</italic>), but not at the locative PP region (<italic>en la mochila</italic>) (1,143 vs. 1,453&#x2009;ms). By contrast, overt pronouns were read significantly faster with an object-antecedent bias (1,308&#x2009;ms) than with a subject-antecedent bias (1,402&#x2009;ms) at the PP region, but not at the object region (884 vs. 887&#x2009;ms). Findings are in line with Carminatti&#x2019;s PAS prediction, though note that (i) RT differences<xref ref-type="fn" rid="fn0004"><sup>4</sup></xref> for overt pronouns (170&#x2009;ms) are smaller than for null pronouns (399&#x2009;ms), which suggests again a rather &#x201C;flexible&#x201D; antecedent bias for overt pronouns; (ii) RT differences are more observable in some regions than in others, which suggests that these stimuli are not straightforwardly parsed probably due to the complex disambiguation mechanism. Further results from adult L1 Arabic-L2 Spanish and L1 English-L2 Spanish learners at three proficiency levels (intermediate, upper intermediate, high) revealed that the advanced learners can eventually parse PAS structures in a native-like fashion, irrespective of their L1 (a (non)null-subject language like English or Arabic).</p>
<p>Intrasentential (subordinate-main order) PAS was investigated in adult Mexican Spanish monolinguals (clause-by-clause presentation) (<xref ref-type="bibr" rid="ref21">Keating et al., 2016</xref>). The ambiguous anaphor is resolved postverbally via world knowledge: <italic>su culpabilidad</italic> &#x201C;his guilt&#x201D; forces a subject antecedent (<italic>el sospechoso</italic> &#x201C;the suspect&#x201D;) in (12a), but an object antecedent in (12b). Null-pronoun clauses were read significantly faster with subject (2,186&#x2009;ms) than with object (2,447&#x2009;ms) antecedents. By contrast, overt-pronoun sentences were read faster with object (2,456&#x2009;ms) than with subject (2,605&#x2009;ms) antecedents. These results confirm Carminatti&#x2019;s PAS but note that if we calculate the RT differences,<xref ref-type="fn" rid="fn0005"><sup>5</sup></xref> the mathematical difference is smaller for overt pronouns (194&#x2009;ms) than for null pronouns (261), which suggests again a certain &#x201C;flexibility&#x201D; for overt pronouns.</p>
<list list-type="simple">
<list-item>
<p>(12) a. Despu&#x00E9;s de que el sospechoso<sub>i</sub> habl&#x00F3; con el polic&#x00ED;a<sub>j</sub>, <inline-formula>
<mml:math id="M14">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi>&#x00E9;</mml:mi>
<mml:msub>
<mml:mi mathvariant="normal">l</mml:mi>
<mml:mi mathvariant="normal">i</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> admiti&#x00F3; su culpabilidad.</p>
</list-item>
<list-item>
<p>b. Despu&#x00E9;s de que el polic&#x00ED;a<sub>i</sub> habl&#x00F3; con el sospechoso<sub>j</sub>, <inline-formula>
<mml:math id="M15">
<mml:mrow>
<mml:mrow>
<mml:mo>{</mml:mo>
<mml:mrow>
<mml:mtable>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:msub>
<mml:mi mathvariant="normal">&#x00D8;</mml:mi>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mrow>
<mml:mi>&#x00E9;</mml:mi>
<mml:msub>
<mml:mi mathvariant="normal">l</mml:mi>
<mml:mi mathvariant="normal">j</mml:mi>
</mml:msub>
</mml:mrow>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:mrow>
<mml:mo>}</mml:mo>
</mml:mrow>
</mml:mrow>
</mml:math>
</inline-formula> admiti&#x00F3; su culpabilidad.</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;{After the suspect talked to the policeman | After the policeman talked to the suspect}, (he) admitted his guilt.&#x201D;</p>
</disp-quote>
<p>In a SPRT, <xref ref-type="bibr" rid="ref33">Mart&#x00ED;n-Villena (2023)</xref> used the same stimuli as in the offline experiment above. Results showed differences depending on the region analyzed. In the subordinate clause segment, null pronouns showed an unclear bias (Subject: 1,383&#x2009;ms; Object: 1,372&#x2009;ms), but overt pronouns exhibited a clear object bias (Subject: 2,129&#x2009;ms; Object 1,940&#x2009;ms). Interestingly, in the comprehension question segment, null pronouns showed a subject bias (Subject: 946&#x2009;ms; Object 1,051&#x2009;ms), but overt pronouns showed an object bias (Subject 1,242&#x2009;ms; Object: 1,037&#x2009;ms), as predicted by the PAS.</p>
</sec>
<sec id="sec5">
<label>1.1.3.</label>
<title>Summary of the experimental evidence: native Spanish</title>
<p>The native Spanish PAS results from the experimental studies are often contradictory. This could be due to multiple factors (many of which were taken into account in our corpus study), e.g.: the different varieties of the monolinguals of Spanish; the PAS configuration (intersentential vs. intrasentential) and the clausal order (main-subordinate vs. subordinate-main); and the different formats (and presentation types) of the offline and online experimental methods, among others.</p>
<p>A visual summary of <italic>offline</italic> PAS biases in native Spanish (<xref ref-type="fig" rid="fig1">Figure 1</xref>) suggests that the original PAS formulation for native Italian is not fully operative in native Spanish: Whereas null pronouns clearly select a subject antecedent (69%&#x2009;~&#x2009;87% range), as predicted by Carminati&#x2019;s PAS, overt pronouns show a &#x201C;flexible&#x201D; preference by often selecting an object antecedent around half of the time (50%&#x2009;~&#x2009;65% range), which implies that the rest of the time they select a subject antecedent. In <italic>online</italic> experiments, null-subject sentences are read significantly faster with forced subject than with forced object antecedents, whereas overt-subject sentences are read faster with forced object than subject antecedents, as predicted by PAS, though note that the subject vs. object RT differences are usually weaker with overt pronouns than with null pronouns, which again suggests a mild &#x201C;flexibility&#x201D; of overt pronouns. The offline and online native Spanish findings thus suggest that, whereas null pronouns have a strong subject bias, overt pronouns are less clear-cut (i.e., more &#x201C;flexible&#x201D;) in their choice of antecedent. We will argue that such flexibility is more apparent than real, as our corpus data will reveal.</p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Summary of offline and online preferences for the PAS in native Spanish and Italian.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g001.tif"/>
</fig>
</sec>
<sec id="sec6">
<label>1.1.4.</label>
<title>Corpus evidence</title>
<p>To our knowledge, there is no corpus-based study targeting specifically the PAS in adult Spanish monolinguals/learners. At best, there is some indirect PAS evidence since the corpus studies to be reviewed analyzed multiple types of AR scenarios (including PAS), so it is unclear to what extent their findings can extrapolate to specific PAS scenarios.</p>
<p>The experimental study reviewed above (<xref ref-type="bibr" rid="ref3">Bel et al., 2016a</xref>) presents additional evidence from a written and spoken production task by Peninsular Spanish monolinguals. The researchers analyzed different types of AR scenarios, including PAS-like scenarios. Null pronouns clearly biased toward subject antecedents (77.27%), while overt pronouns showed a less clear-cut antecedent bias (subject: 42.86%; non-subject: 57.14%). Moroccan Arabic/Spanish early bilinguals&#x2019; null pronouns clearly biased toward subject antecedents (70.19%), while their overt pronouns biased toward both non-subject (35.71%) and subject (64.29%) antecedents. Overt pronouns reflect again the already reported &#x201C;flexibility&#x201D;. Importantly, this study (i) does not report the production of NP anaphors, which are crucial for our understanding of AR in general and the PAS in particular, as we will later show in this paper; (ii) analyses both singular and plural anaphoric forms together, though corpus data has shown that only 3rd singular anaphors are problematic for learners (<xref ref-type="bibr" rid="ref25">Lozano, 2009</xref>, <xref ref-type="bibr" rid="ref26">2016</xref>); and (iii) presents data from teenage Spanish monolinguals <xref ref-type="fn" rid="fn0006"><sup>6</sup></xref> and early bilinguals, so the evidence about how the PAS operates in adult monolinguals and L2 Spanish is rather indirect. In a follow-up study, <xref ref-type="bibr" rid="ref15">Garc&#x00ED;a-Alcaraz and Bel (2019)</xref> used the same task and coding criteria. This time, the Spanish monolinguals were university students, and the L1 Moroccan Arabic-L2 Spanish learners were teenage sequential bilinguals. Results suggest that both monolinguals and bilinguals produce null pronominal subjects to mark topic continuity around 2/3 of the time and topic shift around 1/3. Regarding overt pronouns, their production was very low (4 tokens or less depending on the configuration), which is not very informative. In short, while suggestive, these findings do not fully inform about PAS scenarios in either native or L2 adult Spanish.</p>
<p>A series of corpus studies (<xref ref-type="bibr" rid="ref25">Lozano, 2009</xref>, <xref ref-type="bibr" rid="ref26">2016</xref>; <xref ref-type="bibr" rid="ref34">Mart&#x00ED;n-Villena and Lozano, 2020</xref>) targeted AR scenarios with subject anaphoric forms (null/overt pronouns, as well as NPs). Results from adult Peninsular Spanish monolinguals reveal some consistent findings across studies: whereas null pronouns clearly encode topic continuity, it is NPs that encode topic shift more often than overt pronouns do, particularly when there are several potential antecedents in discourse. L1 English-L2 Spanish learners do not typically show problems in topic-shift contexts (as they use overt forms to avoid ambiguity) but are redundant in topic-continuity contexts (as they overuse overt pronouns). These findings are captured by the Pragmatic Principles Violation Hypothesis (PPVH) (<xref ref-type="bibr" rid="ref26">Lozano, 2016</xref>), which postulates differential effects at the syntax-discourse interface with AR: learners obey the pragmatic Principle of Clarity as they use full anaphoric forms in cases of ambiguity, but they are lax with the Principle of Economy, as they redundantly produce overt anaphoric forms when not required in topic continuity, though can be modulated by the amount of potential antecedents. In short, learners are more redundant than ambiguous. We will get back to the PPVH when discussing our results.</p>
<p>To summarize, the corpus-based findings are clearly insufficient since they: (i) do not specifically target PAS scenarios but rather conflate different types of AR scenarios in their analyses; (ii) some of them do not consider the role of subject NPs as an anaphoric form in its own right. This, coupled by certain limitations in the experimental studies, motivated the formulation of our research questions with a view to answering some unresolved issues in the production of PAS in native and L2 Spanish.</p>
</sec>
</sec>
<sec id="sec7">
<label>1.2.</label>
<title>The current study: research questions and hypotheses</title>
<p>The bulk of experimental studies on AR have investigated the PAS with two potential antecedents (subject/non-subject) and two anaphoric forms (overt/null pronominal subject) in either inter- or intra-sentential configurations. So, what we know about the PAS comes mostly from a series of similarly-designed experiments that do not question whether (i) the PAS may represent an oversimplified way of resolving anaphora in native (and L2) Spanish; (ii) PAS scenarios may be more complex than traditionally assumed (i.e., they can contain more than two antecedents in other syntactic positions); (iii) the antecedents may be realized by other forms other than null/overt pronouns (i.e., NPs for example). Unlike experiments, corpus data can shed light on these questions since they contain natural language production (where AR configurations are neither controlled nor constrained) and offer contextually rich scenarios with anaphors and antecedents embedded in their entire discourse. Unlike experiments, corpus data can shed light on these questions since they (i) contain natural language production where AR configurations are neither controlled nor constrained; (ii) offer contextually rich scenarios with anaphors and antecedents embedded in their entire discourse. This led to <italic>RQ1a</italic> and <italic>RQ1b</italic>.</p>
<disp-quote>
<p><italic>RQ1a</italic> (<italic>Prototypicality of PAS</italic>): Is the PAS a prototypical way of resolving anaphora in native (and in non-native) Spanish, as implicitly assumed in the literature?</p>
</disp-quote>
<disp-quote>
<p><italic>H1a</italic>: The PAS is but one of many possible mechanisms for resolving anaphors in native and non-native Spanish.</p>
</disp-quote>
<disp-quote>
<p><italic>RQ1b</italic> (<italic>Complexity of PAS</italic>): Can the standard PAS configuration (subject/non-subject antecedent; null/overt pronominal subject anaphor) be more complex than assumed in the literature?</p>
</disp-quote>
<disp-quote>
<p><italic>H1b</italic>: Corpus data will reveal that the PAS is richer than standardly assumed, in terms of antecedent configurations, syntactic possibilities and range of anaphoric forms.</p>
</disp-quote>
<p>Experimental PAS studies have typically restricted their focus to two anaphoric forms (overt/null pronominal subjects). Corpus studies have reported the use of other anaphoric forms (e.g., repeated Ns and NPs) in several AR scenarios, so NPs may be also possible Refererential Expression (RE) forms in PAS.<xref ref-type="fn" rid="fn0007"><sup>7</sup></xref></p>
<disp-quote>
<p><italic>RQ2</italic> (<italic>RE forms in discourse</italic>): Apart from null/overt pronominal subjects, are other RE forms possible in native and L2 Spanish PAS?</p>
</disp-quote>
<disp-quote>
<p><italic>H2</italic>: In line with corpus findings on AR in general, we predict for PAS (i) null pronouns to be abundant due to the null-subject nature of Spanish; (ii) overt pronouns to be infrequent and, (ii) importantly, NPs to be more frequent than overt pronouns. The range of REs in PAS will therefore include null/overt pronominal anaphors and NPs (used with an anaphoric value).</p>
</disp-quote>
<p>Experimental studies report Spanish null pronouns to bias toward a preverbal subject antecedent, whereas overt pronouns show a more &#x201C;flexible&#x201D; behavior. This contrasts with native Italian where overt/null pronouns show a clear division of labor. Additionally, experimental studies have not typically included NPs as a possible RE form.</p>
<disp-quote>
<p><italic>RQ3</italic> (<italic>Division of labor</italic>): Regarding the division of labor in native and L2 Spanish, will the &#x201C;flexible&#x201D; behavior of overt pronouns be better accounted for if NPs are also included as a possible type of RE?</p>
</disp-quote>
<disp-quote>
<p><italic>H3</italic>: Null pronouns will be clearly biased toward a subject antecedent, as previously reported, whereas overtly realized REs (i.e., overt pronouns and NPs together) will be clearly biased toward non-subject antecedents. Learners will show growing sensitivity to such division of labor as proficiency increases, but native-like ultimate attainment is not expected for upper-advanced learners since the PAS is constrained at the syntax-discourse interface (cf. <italic>RQ4</italic> below), which is a problematic area for L2 learners (<xref ref-type="bibr" rid="ref28">Lozano, 2021a</xref> for an overview).</p>
</disp-quote>
<p>The implicit assumption in the experimental literature is that purely configurational factors (null&#x2794;subject vs. overt&#x2794;non-subject) overlap with discursive information-status factors (null&#x2794;topic continuity vs. overt&#x2794;topic shift). <italic>RQ4</italic>/<italic>H4</italic> (when contrasted to <italic>RQ3</italic>/<italic>H3</italic>) will determine the extent to which the overlap assumption is correct. This motivates theoretical questions having to do with likely deficits at the syntax-discourse interface.</p>
<disp-quote>
<p><italic>RQ4</italic> (<italic>Syntax-discourse interface</italic>): Will syntactic configuration overlap with information status in PAS configurations and, if so, will learners be eventually (un) able to acquire this syntax-discourse phenomenon?</p>
</disp-quote>
<disp-quote>
<p><italic>H4</italic>: Syntactic configuration will overlap with information status and NPs will play a role (null&#x2794;subject/topic continuity; overt &#x0026; NP&#x2794;non-subject/topic shift). Learners will show an increasing trend toward the native norm, yet the syntax-discourse properties of the PAS will not be fully acquired, as predicted by models like the IH and the PPVH.</p>
</disp-quote>
<p>Despite English being a non-null subject language, corpus data (<xref ref-type="bibr" rid="ref42">Quesada and Lozano, 2020</xref>) have shown that English monolinguals allow null pronouns in very specific contexts: topic continuity and coordination at around 77% (e.g., <italic>Lucy<sub>i</sub> walked for an hour and &#x00D8;<sub>i</sub> had a picnic</italic>), but never in non-coordinate configurations. So, it could be argued that L2 Spanish learners&#x2019; production of null pronouns in topic continuity could be due to L1 transfer rather than actual acquisition, which leads to the following exploratory research question.</p>
<disp-quote>
<p><italic>RQ5</italic> (<italic>Cross-linguistic influence</italic>): Will L2 Spanish learners&#x2019; distribution of null pronouns be a reflection of their allowance in their L1 English (topic continuity and coordination) or will it be a reflection of acquisition at the syntax-discourse interface? It may be the case that learners transfer in initial stages but progressively acquire the discursive distribution of null pronouns.</p>
</disp-quote>
<disp-quote>
<p><italic>H5</italic>: (<italic>Transfer account)</italic></p>
</disp-quote>
<disp-quote>
<p>If L2 Spanish learners are transferring from their L1 English, null subjects will be produced mainly where they are allowed in English (topic continuity with coordination) and not where they are not allowed (topic continuity with non-coordination).</p>
</disp-quote>
<disp-quote>
<p>(<italic>Non-transfer account</italic>, i.e., acquisition account)</p>
</disp-quote>
<disp-quote>
<p>If they are rather sensitive to the pragmatics of null pronouns in Spanish, null subjects will be produced where they are allowed in native Spanish, i.e., across the board (both in coordination and non-coordination).</p>
</disp-quote>
<p>Previous PAS experimental studies are often contradictory depending on the sentential configuration: inter- vs. intra-sentential; main-subordinate vs. subordinate-main orders (<italic>cf.</italic> the tables in the <xref ref-type="supplementary-material" rid="SM1">online</xref> <xref ref-type="supplementary-material" rid="SM1">Supplementary material</xref>). <italic>RQ6</italic> is an exploratory question to explore whether the sentential PAS configuration modulates the choice of RE in naturalistic corpus production.</p>
<disp-quote>
<p><italic>RQ6</italic> (<italic>Sentential configurations</italic>): In which sentential configurations (intra- vs. inter-sentential) will PAS structures be more frequent in naturalistic corpus production? Which PAS clausal order (main-subordinate vs. subordinate-main) is prototypical? Will learners&#x2019; production ultimately approach to/deviate from Spanish monolinguals?</p>
</disp-quote>
</sec>
</sec>
<sec sec-type="methods" id="sec8">
<label>2.</label>
<title>Method</title>
<sec id="sec9">
<label>2.1.</label>
<title>Corpus: CEDEL2</title>
<p><italic>Corpus Escrito del Espa&#x00F1;ol L2</italic> (CEDEL2) (<xref ref-type="bibr" rid="ref30">Lozano, 2022</xref>) is a multi-L1 corpus of L2 Spanish learners coming from 11 different L1 backgrounds, plus a Spanish monolingual control subcorpus. CEDEL2 (version 2) currently holds 1,105,936 words, 4,399 participants, and 14 task topics. It is freely available/downloadable at <ext-link xlink:href="http://cedel2.learnercorpora.com/" ext-link-type="uri">http://cedel2.learnercorpora.com</ext-link>.</p>
<p>Data are collected via online forms<xref ref-type="fn" rid="fn0008"><sup>8</sup></xref> and participants complete three forms: (i) linguistic background; (ii) standardized placement test (just for learners) (<xref ref-type="bibr" rid="ref45">University of Wisconsin, 1998</xref>); and (iii) written/spoken text.</p>
</sec>
<sec id="sec10">
<label>2.2.</label>
<title>Sample</title>
<p>We selected an L1 English-L2 Spanish (plus a comparable Spanish monolingual control) sample (<xref ref-type="table" rid="tab1">Table 1</xref>) based on the following criteria: (i) the participant&#x2019;s age range was 18&#x2009;~&#x2009;40, since Working Memory, which may affect AR, appears to decay after the age of 40 (<xref ref-type="bibr" rid="ref4">Bel et al., 2016b</xref>); (ii) learners&#x2019; proficiency-level range was intermediate~advanced; and (iii) only two composition titles were targeted (<italic>cf.</italic> 2.3 below).<xref ref-type="fn" rid="fn0009"><sup>9</sup></xref> Two hundred two texts met these criteria but we finally selected those that had at least one instance of a PAS (<italic>N</italic>&#x2009;=&#x2009;75). We originally departed from two intermediate groups: lower intermediates (placement score: 21&#x2009;~&#x2009;28 raw score, 49%&#x2009;~&#x2009;65%) and upper intermediates (29&#x2009;~&#x2009;35, 67&#x2013;81%). Since they did not significantly differ in our analyses, we decided to analyze both groups as a single group of intermediates to simplify the between-group statistical analyses and interpretations. Learners had an equivalent age of exposure (AoE) to L2 Spanish and their length of instruction (LoI) in Spanish and length of stay abroad (LoSA) in a Spanish-speaking country increased with proficiency.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>Texts and participants&#x2019; bio-data.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top" rowspan="2">Group</th>
<th align="center" valign="top" colspan="2">Intermediate</th>
<th align="center" valign="top" rowspan="2">Lower advanced</th>
<th align="center" valign="top" rowspan="2">Upper advanced</th>
<th align="center" valign="top" rowspan="2">Monolinguals</th>
</tr>
<tr>
<th align="center" valign="top">Lower</th>
<th align="center" valign="top">Upper</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">Placement raw score (0&#x2013;43)</td>
<td align="char" valign="top" char=".">21&#x2009;~&#x2009;28</td>
<td align="char" valign="top" char=".">29&#x2009;~&#x2009;35</td>
<td align="char" valign="top" char=".">36&#x2009;~&#x2009;40</td>
<td align="char" valign="top" char=".">41&#x2009;~&#x2009;43</td>
<td align="char" valign="top" char=".">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">Equivalent percentage (0&#x2013;100%)</td>
<td align="char" valign="top" char=".">49&#x2009;~&#x2009;65%</td>
<td align="char" valign="top" char=".">66&#x2009;~&#x2009;81%</td>
<td align="char" valign="top" char=".">82&#x2009;~&#x2009;94%</td>
<td align="char" valign="top" char=".">95&#x2009;~&#x2009;100%</td>
<td align="char" valign="top" char=".">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">Texts that met criteria</td>
<td align="char" valign="top" char="." colspan="2">69</td>
<td align="char" valign="top" char=".">37</td>
<td align="char" valign="top" char=".">13</td>
<td align="char" valign="top" char=".">103</td>
</tr>
<tr>
<td align="left" valign="middle">Texts analyzed</td>
<td align="char" valign="top" char="." colspan="2">21</td>
<td align="char" valign="top" char=".">19</td>
<td align="char" valign="top" char=".">8</td>
<td align="char" valign="top" char=".">27</td>
</tr>
<tr>
<td align="left" valign="middle">Mean age</td>
<td align="char" valign="top" char="." colspan="2">20.8</td>
<td align="char" valign="top" char=".">21.3</td>
<td align="char" valign="top" char=".">25.5</td>
<td align="char" valign="top" char=".">25.6</td>
</tr>
<tr>
<td align="left" valign="middle">Mean proficiency</td>
<td align="char" valign="top" char="." colspan="2">69.6%</td>
<td align="char" valign="top" char=".">86.8%</td>
<td align="char" valign="top" char=".">96.7%</td>
<td align="char" valign="top" char=".">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">AoE (years)</td>
<td align="char" valign="top" char="." colspan="2">14</td>
<td align="char" valign="top" char=".">14.5</td>
<td align="char" valign="top" char=".">12.2</td>
<td align="char" valign="top" char=".">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">LoI (years)</td>
<td align="char" valign="top" char="." colspan="2">5.3</td>
<td align="char" valign="top" char=".">5.5</td>
<td align="char" valign="top" char=".">8.8</td>
<td align="char" valign="top" char=".">&#x2013;</td>
</tr>
<tr>
<td align="left" valign="middle">LoSA (months)</td>
<td align="char" valign="top" char="." colspan="2">7.4</td>
<td align="char" valign="top" char=".">11.1</td>
<td align="char" valign="top" char=".">12.3</td>
<td align="char" valign="top" char=".">&#x2013;</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="sec11">
<label>2.3.</label>
<title>Task</title>
<p>We selected two task tittles (<italic>Talk about a famous person</italic> and <italic>Summarize a film you have recently seen</italic>), since they are narratives that contain (i) abundant [+human] 3<sup>rd</sup> person antecedent-anaphor chains; and (ii) PAS constructions, which were more frequent in the second task than in the first task and which offered different characters in discourse suitable for the topic continuity/shift purpose of this study.</p>
</sec>
<sec id="sec12">
<label>2.4.</label>
<title>Corpus annotation and tagset</title>
<p>The corpus sample was manually annotated (i.e., tagged) with UAM Corpus Tool (<xref ref-type="bibr" rid="ref37">O&#x2019;Donnell, 2009</xref>), version 6.2j (February 2023).<xref ref-type="fn" rid="fn0010"><sup>10</sup></xref> We firstly tagged each text to indicate the group category (intermediate, lower advanced, upper advanced, and monolingual), which allows between-group comparisons for the same linguistic feature, as will be explained below. We designed another tagset to count the frequency of two AR scenarios (PAS vs. other AR). Each RE in subject position was assigned either the <italic>PAS</italic> tag (when the RE was preceded by a subject/non-subject antecedents) or <italic>other</italic> (when the RE was preceded by an AR scenario other than PAS). <xref ref-type="fig" rid="fig2">Figure 2</xref> shows the fine-grained, linguistically-informed tagset to annotate PAS.<xref ref-type="fn" rid="fn0011"><sup>11</sup></xref> It allows for multiple and intricate statistical analyses among tags, as will become obvious later. It is inspired by previous corpus studies on AR (<xref ref-type="bibr" rid="ref26">Lozano, 2016</xref>; <xref ref-type="bibr" rid="ref42">Quesada and Lozano, 2020</xref>), although we introduced new features.</p>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>PAS annotation tagset.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g002.tif"/>
</fig>
<p>Every 3rd person human subject that followed the syntactic configuration of the PAS was manually tagged. First, the <bold>PAS-type</bold> system included: (i) <italic>standard</italic> PAS with two antecedents, as in (13), and (ii) <italic>complex</italic> PAS with more than two antecedents, as in (14a-c). For example, the tag used to annotate complex PAS in (14a) is <italic>s1_nons2_nons3</italic>, which indicates we have 3 potential antecedents: the first one is in subject position (s1) and the other two in non-subject position realized via a complex NP: PP (nons3) within an NP (nons2). In (14b), the tag <italic>s1_nons2andnons3</italic> indicates that there is a singular antecedent in subject position (s1, which happens to be a null pronouns) followed by two NP coordinated antecedents in non-subject position (s2&#x0026;s3) embedded within a PP. Notice that, due to the complexity of the antecedents&#x2019; region, the anaphor is a complex NP for disambiguation purposes. Other complex PAS contained plural REs, as in (14c), but we excluded them from our current analysis since it has been shown that the truly problematic cases of AR are 3rd person singular and not plural (<xref ref-type="bibr" rid="ref25">Lozano, 2009</xref>).</p>
<list list-type="simple">
<list-item>
<p>(13) Standard PAS:</p>
</list-item>
</list>
<p><italic>Naaven<sub>i</sub></italic> se ha enamorado de <italic>Tiana<sub>j</sub></italic> y <bold>&#x00D8;</bold><sub>
<bold>i</bold>
</sub> quiere pedirle matrimonio. [Monolingual: ES_WR_24_3_IZG.txt].<xref ref-type="fn" rid="fn0012"><sup>12</sup></xref></p>
<disp-quote>
<p>&#x201C;Naaven<sub>i</sub> has fallen in love with Tiana<sub>j</sub> and &#x00D8;<sub>i</sub> wants to propose to her&#x201D;.</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>(14) Complex PAS:</p>
</list-item>
</list>
<p>a.<italic>La chica<sub>i</sub> se</italic> enamora del <italic>amante<sub>j</sub></italic> de su <italic>madre<italic>k</italic></italic> hasta que al final &#x00D8;<sub>
<bold>i</bold>
</sub> acaba teniendo &#x2026; [Monolingual: ES_WR_30_3_JVM].</p>
<disp-quote>
<p>&#x201C;The girl<sub>i</sub> falls in love with the lover<sub>j</sub> of her mother<sub>k</sub> until &#x00D8;<sub>i</sub> ends up having&#x2026;&#x201D;.</p>
</disp-quote>
<list list-type="alpha-lower">
<list-item>
<p>Pero el principal problema que <bold>&#x00D8;</bold><sub>
<bold>i</bold>
</sub> ten&#x00ED;a era que <bold>&#x00D8;</bold><sub>
<bold>i</bold>
</sub> sufr&#x00ED;a un maltrato constante por parte de su <bold>madre</bold><sub>
<bold>j</bold>
</sub> y del <bold>novio</bold><sub>
<bold>k</bold>
</sub> de <bold>&#x00E9;sta</bold><sub>
<bold>j</bold>
</sub>. <bold>El novio</bold><sub>
<bold>k</bold>
</sub> <bold>de la madre</bold><sub>
<bold>j</bold>
</sub> hab&#x00ED;a&#x2026; [Monolingual: ES_WR_31_3_EAC]</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;But the main problem &#x00D8;<sub>i</sub> had was that &#x00D8;<sub>i</sub> was abused by her mother<sub>j</sub> and the boyfriend<sub>k</sub> of her<sub>j</sub>. The boyfriend<sub>k</sub> of the mother<sub>j</sub> had&#x2026;)&#x201D;.</p>
</disp-quote>
<list list-type="alpha-lower">
<list-item>
<p><italic>&#x00D8;<sub>ij</sub></italic> Juntos tendr&#x00E1;n que huir de <italic>Dr. Facilier<sub>k</sub></italic> a los pantanos, dnde <bold>&#x00D8;</bold><sub>
<bold>ij</bold>
</sub> se encuentran&#x2026; [Monolingual: ES_WR_24_3_IZG]</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;&#x00D8;<sub>ij</sub> Together will have to escape from Dr. Facilier<sub>k</sub> to the swamps, where <bold>&#x00D8;</bold><sub>
<bold>ij</bold>
</sub> meet &#x2026;&#x201D;.</p>
</disp-quote>
<p>The <bold>anaphor-form</bold> system includes the RE form (null/overt pronouns and NPs) in subject position, as shown in bold in (15). The <bold>anaphor-number</bold> system includes the RE number (singular/plural), which served us to exclude plural REs in the analyses, as justified above.</p>
<list list-type="simple">
<list-item>
<p>(15) &#x2026; <bold>el protagonista</bold><sub>i</sub> de la pel&#x00ED;cula se enamora de la chica<sub>j</sub> y <bold>ella</bold><sub>j</sub> le<sub>i</sub> pide por favor que <bold>&#x00D8;</bold><sub>i</sub> deje el negocio &#x2026; [Monolingual: ES_WR_23_3_EM].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;&#x2026;the <bold>main character</bold><sub>i</sub> of the film falls in love with the girl<sub>j</sub> and <bold>she</bold><sub>j</sub> asks him<sub>i</sub> that <bold>&#x00D8;</bold><sub>i</sub> leaves the business &#x2026;&#x201D;</p>
</disp-quote>
<p>The <bold>information-status</bold> system comprises topic-continuity and topic-shift contexts, as in (16a-b) respectively. The <bold>antecedent-function</bold> system included subject antecedent, non-subject antecedent, and subject/non-subject antecedent (for cases of complex PAS). This system allowed us to detect PAS scenarios with subject-antecedent biases, as in (16a), or non-subject antecedent biases, as in (16b).</p>
<list list-type="simple">
<list-item>
<p>(16) a. Un periodista<sub>i</sub> investiga la desaparici&#x00F3;n de una rica heredera<sub>j</sub>, hace cuarenta a&#x00F1;os. Para ello, <bold>&#x00D8;</bold><sub>i</sub> cuenta con&#x2026;</p>
<p>[Monolingual: ES_WR_24_3_AW].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;A journalist<sub>i</sub> investigates the disappearance of a rich heiress<sub>i</sub>, 40 years ago. To do so, <bold>&#x00D8;</bold><sub>i</sub> relies on&#x2026;&#x201D;</p>
</disp-quote>
<p>b. Bella<sub>i</sub> se da cuenta de que Jacob<sub>j</sub> est&#x00E1; enamorado de ella<sub>i</sub> y <bold>ella</bold><sub>i</sub> tambi&#x00E9;n un poco de &#x00E9;l<sub>j</sub> [Monolingual: ES_WR_21_3_ICH].</p>
<disp-quote>
<p>&#x201C;Bella<sub>i</sub> realizes that Jacob<sub>j</sub> is in love with her<sub>i</sub> and <bold>she</bold><sub>i</sub> is also in love with him<sub>j</sub>&#x201D;.</p>
</disp-quote>
<p>In the <bold>syntactic-configuration</bold>, we tagged the type of intra-sentential and inter-sentential configurations, e.g., topic-continuity and coordination in (16b) and topic continuity and non-coordination, which can be of different types, e.g., subordination in (17) or new sentence in (18).</p>
<list list-type="simple">
<list-item>
<p>(17) &#x2026; un padre<sub>i</sub> trata por todos los medios de llevar a su hijo<sub>j</sub> de diez a&#x00F1;os hasta el mar, donde <bold>&#x00D8;</bold><sub>i</sub> espera encontrar&#x2026; [Monolingual: ES_WR_22_3_AFL].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;&#x2026;a father<sub>i</sub> tries by all means to take his ten-year-old son<sub>j</sub> to the sea, where <bold>&#x00D8;</bold><sub>i</sub> hopes to find &#x2026;&#x201D;</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>(18) &#x2026; y &#x00D8;<sub>i</sub> llega a cortarle<sub>j</sub> un dedo de un hachazo. Despu&#x00E9;s <bold>&#x00D8;</bold><sub>i</sub> intenta matar a George<sub>k</sub>&#x2026; [Monolingual: ES_WR_28_3_MAAO].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;&#x2026; and &#x00D8;<sub>i</sub> cuts off her<sub>j</sub> finger with an axe. Later &#x00D8;<sub>i</sub> tries to kill George<sub>k</sub>&#x2026;&#x201D;.</p>
</disp-quote>
<p>Finally, the <bold>anaphora-resolution</bold> system indicates the type of resolution: via morphosyntax or semantics. In this paper, we analyzed only the PAS that was morphosyntactically resolved. In order to avoid skewing our results, we excluded PAS that was semantically resolved (i.e., null pronouns in topic-shift scenarios like (19), which are ultimately resolved via directive verbs).</p>
<list list-type="simple">
<list-item>
<p>(19) Ella<sub>i</sub> le<sub>j</sub> pide que <bold>&#x00D8;</bold><sub>j</sub> espere a&#x2026;[Monolingual: ES_WR_26_3_MPVI].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;She<sub>i</sub> asks him<sub>j</sub> that <bold>&#x00D8;</bold><sub>j</sub> waits for her<sub>i</sub> to&#x2026;&#x201D;.</p>
</disp-quote>
</sec>
<sec id="sec13">
<label>2.5.</label>
<title>Analysis</title>
<p>UAM Corpus Tool has an in-built statistical analysis software. Between-group (or between-system/tag) comparisons are based on the tags&#x2019; raw frequencies and statistical contrasts are chi square (<italic>&#x03C7;</italic><sup>2</sup>) tests, accompanied by their significance level (<italic>p</italic>) and their effect size (Cohen&#x2019;s <italic>h</italic>).</p>
<p>Based on the linguistically-motivated tagging scheme (<xref ref-type="fig" rid="fig2">Figure 2</xref>), UAM Corpus Tool allows for multiple and sophisticated statistical contrasts between the different groups and the (sub) nodes and terminal nodes of the tagset. These contrasts were motivated by the linguistically-informed hypotheses from section 1.2. Following statistical recommendations for corpus data (<xref ref-type="bibr" rid="ref9">Egbert et al., 2020</xref>), we purposely decided to use the <italic>&#x03C7;</italic><sup>2</sup> statistical contrasts provided by the software rather than submitting the data to more sophisticated statistical analyses (which involve transforming the data and abstracting away from the linguistic facts and interpretations):</p>
<disp-quote>
<p>&#x201C;the most appropriate method for the task at hand should not be the most sophisticated method &#x2026; Instead, we should always strive to choose minimally sufficient statistical methods, meaning that we should choose tests that are no more nor less sophisticated than the study design requires. The reason for this is twofold: (1) all descriptive and inferential statistical tests force us to abstract away from language to some extent and (2) there is often an inverse relationship between the level of sophistication of the method and the linguistic interpretability of the results.&#x201D; (<xref ref-type="bibr" rid="ref9">Egbert et al., 2020</xref>, p. 40)</p>
</disp-quote>
</sec>
</sec>
<sec sec-type="results|discussion" id="sec14">
<label>3.</label>
<title>Results and discussion</title>
<p>We next present and discuss the results for each research question. We leave the general discussion for section 4.</p>
<sec id="sec15">
<label>3.1.</label>
<title><italic>RQ1</italic>/<italic>H1</italic>: frequency of PAS scenarios in natural language production</title>
<p>In <xref ref-type="fig" rid="fig3">Figure 3</xref>, PAS scenarios (gray bars) were compared against other types of AR scenarios (black bars). Spanish monolinguals resolve anaphora via scenarios (68.2%, i.e., 296 REs out of a total of 434 tagged REs) other than standard PAS (21.2%) or complex PAS (10.6%). Thus, standard PAS only amounts to around 1/5th of the total possible AR scenarios. Learners show a similar pattern to monolinguals across all proficiency levels, though only the upper advanced group shows native-like behavior (standard and complex PAS <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;2.16, <italic>p</italic>&#x2009;=&#x2009;0.1419 n.s., <italic>h</italic>&#x2009;=&#x2009;0.204; other scenarios <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;0.15, <italic>p</italic>&#x2009;=&#x2009;0.6964 n.s, <italic>h</italic>&#x2009;=&#x2009;0.030). The lower-level learner groups significantly differ from Spanish monolinguals in other scenarios but not in standard and complex PAS scenarios (intermediates vs. monolinguals: standard and complex PAS <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;0.49, <italic>p</italic>&#x2009;=&#x2009;0.4848 n.s., <italic>h</italic>&#x2009;=&#x2009;0.087, other scenarios <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;8.80 <italic>p</italic>&#x2009;=&#x2009;0.0030, <italic>h</italic>&#x2009;=&#x2009;0.193; lower-advanced vs. monolinguals: standard and complex PAS <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;0.66, <italic>p</italic>&#x2009;=&#x2009;0.4180 n.s., <italic>h</italic>&#x2009;=&#x2009;0.106, other scenarios <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;12.44, <italic>p</italic>&#x2009;=&#x2009;0.0004, <italic>h</italic>&#x2009;=&#x2009;0.235).</p>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>AR scenarios by group.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g003.tif"/>
</fig>
<p>Our findings support <italic>H1a</italic> (PAS represents one of the many possible mechanisms of AR in native and non-native Spanish) and <italic>H1b</italic> (PAS can contain more complex configurations than those traditionally reported in the literature). Corpus data reveal that the traditional assumption of standard PAS as a prototypical strategy to resolve anaphora has been overestimated in the experimental literature.</p>
</sec>
<sec id="sec16">
<label>3.2.</label>
<title><italic>RQ2</italic>/<italic>H2</italic>: overall use of REs in PAS scenarios</title>
<p><italic>RQ2</italic> explores the different RE forms in PAS scenarios, independently from the factors that constrain their choice. Spanish monolinguals produced mostly null pronominal subjects (66.1%), followed by NPs (23.2%) and overt pronominal subjects (10.7%) (<xref ref-type="fig" rid="fig4">Figure 4</xref>). Learners show a tendency toward the native norm as proficiency increases, yet only upper-advanced leaners (57.9% null, 26.3 overt, 15.8% NP) show a rather similar and non-significant pattern to the Spanish monolinguals (null pronouns: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;1.30, <italic>p</italic>&#x2009;=&#x2009;0.2551, <italic>h</italic>&#x2009;=&#x2009;0.169; NPs: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;1.55, <italic>p</italic>&#x2009;=&#x2009;0.2135, <italic>h</italic>&#x2009;=&#x2009;0.188), though a significant difference for overt pronouns (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;7.80, <italic>p</italic>&#x2009;=&#x2009;0.0052, <italic>h</italic>&#x2009;=&#x2009;0.410). The lower-advanced group shows similar proportions for all three RE forms (34.5% null, 35.7% overt, 29.8% NP), which significantly differ from monolinguals for null (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;19.16, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;0.642) and overt (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;17.82, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;0.614), but are non-significant for NPs (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;1.07, <italic>p</italic>&#x2009;=&#x2009;0.3012, <italic>h</italic>&#x2009;=&#x2009;0.149). Intermediates produce mainly overt REs (overt pronouns 39.1%; NPs 38.1%) and some null pronouns (22.8%), with the three RE production rates being significantly different from monolinguals (overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;22.67, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;0.685; NP: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;5.30, <italic>p</italic>&#x2009;=&#x2009;0.0213, <italic>h</italic>&#x2009;=&#x2009;0.324; null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;37.96, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;0.902).</p>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>Overall production of REs across groups.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g004.tif"/>
</fig>
<p>These findings support <italic>H2</italic>. Whereas null pronouns are the tendency in Spanish monolinguals and in upper-advanced learners, the rest of learners differ from monolinguals and show more variability in RE forms. Null pronominal subjects are gradually acquired with proficiency level, whereas overt pronouns show the opposite pattern. Crucially, NPs are a frequent RE form to resolve anaphora in PAS scenarios for both learners and monolinguals. We turn next to the division of labor of such RE forms.</p>
</sec>
<sec id="sec17">
<label>3.3.</label>
<title><italic>RQ3</italic>/<italic>H3</italic>: division of labor of the different anaphoric forms</title>
<p>First, we focus on Spanish monolinguals&#x2019; production to clarify the division of labor in PAS scenarios and to settle the question of whether the alleged flexibility of overt pronouns is more apparent than real. <xref ref-type="fig" rid="fig5">Figure 5</xref> shows a clear bias of null pronouns (93.6%) toward subject antecedents (13), which confirms the PAS and supports most previous research in Spanish. Overt pronouns (32.4%) show a timid bias toward non-subject antecedents, (16b), as previously reported in the literature but, crucially, if we include NPs as a possible RE form, NPs show a strong bias (64.7%) toward non-subject antecedents, (20). Thus, NPs play an important role in PAS scenarios and this could explain the apparently &#x201C;flexible&#x201D; bias found for overt pronouns previously reported.</p>
<list list-type="simple">
<list-item>
<p>(20) &#x00C9;l<sub>i</sub> acaba rechaz&#x00E1;ndola<sub>j</sub> as&#x00ED; que <bold>la chica</bold><sub>
<bold>j</bold>
</sub> harta de&#x2026; [Monolingual: ES_WR_30_3_JVM].</p>
</list-item>
</list>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>Monolinguals&#x2019; production of REs (null/overt/NP) for subject/non-subject antecedents.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g005.tif"/>
</fig>
<disp-quote>
<p>&#x201C;He<sub>i</sub> ends up rejecting her<sub>j</sub>, so <bold>the girl</bold><sub>
<bold>j</bold>
</sub>, being fed up with&#x2026;&#x201D;.</p>
</disp-quote>
<p>Importantly, if we consider overt and NPs forms together (overtly realized REs), then a neater division of labor shows up (<xref ref-type="fig" rid="fig6">Figure 6</xref>): null pronouns are biased toward subject antecedents (93.6%) yet overtly realized REs are biased toward non-subject antecedents (97.1%). Thus, corpus data reveals that the division of labor of AR in native Spanish is more complex and more clear-cut than previously assumed since NPs play a key role. These findings explain the division of labor in native Spanish and therefore settle the dispute on the apparent flexibility of overt pronouns in PAS scenarios.</p>
<fig position="float" id="fig6">
<label>Figure 6</label>
<caption>
<p>Monolinguals&#x2019; production of REs (null vs. overtly realized REs) for subject/non-subject antecedents.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g006.tif"/>
</fig>
<p>Let us now compare learners against monolinguals regarding the production of RE forms for subject vs. non-subject antecedents. As for subject-antecedent biases (<xref ref-type="fig" rid="fig7">Figure 7</xref>), Spanish monolinguals show a clear-cut bias as they produce almost exclusively null pronominal subjects (93.6%). Intermediates show equal variability across all three RE forms (null 35.2%, overt 35.2%, NP 29.6%), as illustrated in (21a, b, c) respectively, and their production is significantly different from monolinguals (null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;44.05, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.358; overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;26.09, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.270; NP: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;10.87, <italic>p</italic>&#x2009;=&#x2009;0.0010, <italic>h</italic>&#x2009;=&#x2009;0.638). From lower advanced to upper advanced we can see an increasing trend toward the native norm, particularly for null pronouns (lower advanced: null 47.8%, NP 26.1%, overt 26.1%; upper advanced: null 75.5%, overt 17.8, NP 6.7%), though, crucially, each advanced group significantly differs from the monolingual group: lower advanced vs. monolinguals (null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;28.75, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.101; overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;18.20, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.072; NP: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;8.07, <italic>p</italic>&#x2009;=&#x2009;0.0045, <italic>h</italic>&#x2009;=&#x2009;0.558); upper advanced vs. monolinguals (null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;7.00, <italic>p</italic>&#x2009;=&#x2009;0.0081, <italic>h</italic>&#x2009;=&#x2009;0.521; overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;11.91, <italic>p</italic>&#x2009;=&#x2009;0.0006, <italic>h</italic>&#x2009;=&#x2009;0.870; except for NPs, where there are no significant differences <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;0.00, <italic>p</italic>&#x2009;=&#x2009;0.9646, <italic>h</italic>&#x2009;=&#x2009;0.009). In short, intermediates know that a null pronoun can select a subject antecedent, but they equally produce overt pronouns (as in their L1) and NPs. Nearly half of the productions of lower-advanced learners are null pronouns. Upper-advanced learners show native-like discriminations, but their productions are not fully native-like yet.</p>
<list list-type="simple">
<list-item>
<p>(21) a. Brooke<sub>i</sub> despide al maestro<sub>j</sub> y <bold>&#x00D8;</bold><sub>
<bold>i</bold>
</sub> emplea Elle<sub>k</sub>&#x2026;[Learner: EN_WR_31_21_7_3_DNP].</p>
</list-item>
</list>
<fig position="float" id="fig7">
<label>Figure 7</label>
<caption>
<p>Production of REs for subject antecedents across groups.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g007.tif"/>
</fig>
<disp-quote>
<p>&#x201C;Brooke<sub>i</sub> fires the teacher<sub>j</sub> and <bold>&#x00D8;</bold><sub>
<bold>i</bold>
</sub> employs Elle<sub>k</sub>&#x2026;&#x201D;</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>b. La madre<sub>i</sub> es sumisa al padre<sub>j</sub> a trav&#x00E9;s de la pel&#x00ED;cula. <bold>Ella</bold><sub>i</sub> no ha sabido&#x2026; [Learner: EN_WR_25_22_17_3_BBB].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;The mother<sub>i</sub> is submissive to the father<sub>j</sub> throughout the film. <bold>She</bold><sub>
<bold>i</bold>
</sub> did not know&#x2026;&#x201D;</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>c. Rose<sub>i</sub> quiere a ve Jack<sub>j</sub> as&#x00ED; que <bold>Rose</bold><sub>i</sub> busca a Jack<sub>j</sub>. [Learner: EN_WR_26_18_3_3_BRS].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Rose<sub>i</sub> wants to see Jack<sub>j</sub> so <bold>Rose</bold><sub>
<bold>i</bold>
</sub> looks for Jack<sub>j</sub>.&#x201D;</p>
</disp-quote>
<p>Consider now non-subject antecedent biases (<xref ref-type="fig" rid="fig8">Figure 8</xref>). Spanish monolinguals&#x2019; production clearly indicates that NPs (64.7%) (and not overt pronouns) are the privileged RE form to refer to a non-subject antecedent. Crucially, null pronouns are hardly an option for any group, so learners know from the outset that a null pronoun is not an adequate form to refer to a non-subject antecedent. It is therefore remarkable that no null pronouns are used in purely structural PAS configurations. As for learners, overt pronouns and NPs are highly produced, but learners are rather indeterminate about them, particularly intermediates, who show optionality in their production (47.2% overt vs. 52.8% NP), and the two advanced groups, who also show a rather indeterminate pattern where overt pronouns are slightly higher than NPs, as in (22 a, b): lower advanced (56.3% vs. 40.6%), upper advanced (54.6% vs. 40.9%).</p>
<list list-type="simple">
<list-item>
<p>(22) a. Bond<sub>i</sub> encuentra Vesper<sub>j</sub>, y ella<sub>j</sub> se disculpa. [Learner: EN_WR_36_19_5_3_MWB].</p>
</list-item>
</list>
<fig position="float" id="fig8">
<label>Figure 8</label>
<caption>
<p>Production of REs for non-subject antecedents across groups.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g008.tif"/>
</fig>
<disp-quote>
<p>&#x201C;Bond<sub>i</sub> finds Vesper<sub>j</sub> and <bold>she</bold><sub>
<bold>j</bold>
</sub> apologizes.&#x201D;</p>
</disp-quote>
<p>b. &#x2026;ella<sub>i</sub> escribe algunas cartas a Michael<sub>j</sub>. Pero Michael<sub>j</sub> no responde. [Learner: EN_WR_38_9_30_3_JG].</p>
<disp-quote>
<p>&#x201C;&#x2026;she<sub>i</sub> writes some letters to Michael<sub>j</sub>. But Michael<sub>j</sub> does not reply.&#x201D;</p>
</disp-quote>
<p><xref ref-type="fig" rid="fig8">Figure 8</xref> visually shows that the learners&#x2019; pattern is either optional (intermediates) or somewhat opposite to the monolinguals&#x2019; (advanced groups). The low frequencies in production in all groups may explain why no significant differences are observed between each of the learner groups and the monolinguals (<italic>p</italic>&#x2009;&#x003E;&#x2009;0.05 in all cases, though <italic>p</italic>&#x2009;&#x003C;&#x2009;0.50 for each of the advanced groups vs. the monolinguals, which represent marginally non-significant differences).</p>
<p>These findings, taken together, support <italic>H3</italic> since null pronouns show a strong bias toward subject antecedents (with learners showing an increasing sensitivity to this), whereas overt material (i.e., overt pronouns as well as NPs) shows a clear bias toward non-subject antecedents.</p>
</sec>
<sec id="sec18">
<label>3.4.</label>
<title><italic>RQ4</italic>/<italic>H4</italic>: the syntax-discourse interface</title>
<p>Recall that, at this point, we need to discriminate between purely structural PAS results (RQ 3, previous section) from purely information status/discursive PAS results (<italic>RQ4</italic>, this section). This will allow us to determine whether the traditional assumption of a correspondence/overlap between syntactic position (subject/non-subject) and information status (topic continuity/shift), as stated in section 1.2, is reflected in production data. Recall that <italic>RQ4</italic> will additionally allow us to check for possible deficits at the syntax-discourse interface, as predicted by the IH.</p>
<p><xref ref-type="fig" rid="fig9">Figure 9</xref> shows the use of REs in topic-continuity contexts, where the production of null pronominal subjects is higher for all groups, although the percentages between groups vary considerably. There is a clear increase of nulls from the intermediate to the monolingual group: intermediate (38.8%), lower-adv (47.8%), upper-adv (76.2%), monolingual (95%). If we compare these results with <xref ref-type="fig" rid="fig7">Figure 7</xref>, we can observe a similar trend in the results and a similar statistical behavior. In particular, intermediates show again similar variability across all three RE forms (null 38.8%, overt, 30.6%, NP 30.6%), as shown in (23a-c) respectively and their production is significantly different from monolinguals (null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;40.39, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.346; overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;21.30, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.173; NP: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;12.83, <italic>p</italic>&#x2009;=&#x2009;0.0003, <italic>h</italic>&#x2009;=&#x2009;0.772). From lower advanced to upper advanced we can see again an increase toward the native norm, particularly for null pronouns (lower advanced: null 47.8%, NP 26.1%, overt 26.1%; upper advanced: null 76.2%, overt 19%, NP 4.8%), though, once again, each advanced group significantly differs from the monolingual group: lower advanced vs. monolinguals (null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;30.52, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.163; overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;17.65, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.072; NP: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;9.53, <italic>p</italic>&#x2009;=&#x2009;0.0020, <italic>h</italic>&#x2009;=&#x2009;0.621); upper advanced vs. monolinguals (null: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;7.86, <italic>p</italic>&#x2009;=&#x2009;0.0050, <italic>h</italic>&#x2009;=&#x2009;0.568; overt: <italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;12.40, <italic>p</italic>&#x2009;=&#x2009;0.0004, <italic>h</italic>&#x2009;=&#x2009;0.903; except for NPs again, where there are no significant differences (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;0.00, <italic>p</italic>&#x2009;=&#x2009;0.9563, <italic>h</italic>&#x2009;=&#x2009;0.011).</p>
<fig position="float" id="fig9">
<label>Figure 9</label>
<caption>
<p>Production of REs in topic-continuity contexts across groups.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g009.tif"/>
</fig>
<list list-type="simple">
<list-item>
<p>(23) a. Rose<sub>i</sub> deja su madre<sub>j</sub> y Cal<sub>k</sub> y &#x00D8;<sub>i</sub> va a buscar Jack<sub>l</sub>. [Learner: EN_WR_26_18_3_3_BRS].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Rose<sub>i</sub> leaves her mother<sub>j</sub> and Cal<sub>k</sub> and <bold>&#x00D8;</bold><sub>
<bold>i</bold>
</sub> goes to find Jack<sub>l.</sub>&#x201D;</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>b. Un d&#x00ED;a el hombre<sub>i</sub> estaba sentado en la selva y &#x00D8;<sub>i</sub> vio la dictadora<sub>j</sub> y despu&#x00E9;s <bold>&#x00E9;l</bold><sub>i</sub> vio un tigre&#x2026;[Learner: EN_WR_31_20_Unknown_STS].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;One day the man<sub>i</sub> was sitting in the jungle and &#x00D8;<sub>i</sub> saw the dictatress<sub>j</sub> and then <bold>he</bold><sub>
<bold>i</bold>
</sub> saw a tiger&#x2026;&#x201D;.</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>c. &#x2026; un hombre<sub>i</sub> muy rico quiere Satine<sub>j</sub>. <bold>El hombre rico</bold><sub>i</sub> tiene mas poder que el hombre pobre<sub>k</sub>. [Learner: EN_WR_35_20_10_3_CES].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;A very rich man<sub>i</sub> loves Satine<sub>j</sub>. <bold>The rich man</bold><sub>i</sub> has more power than the poor man<sub>k</sub>.&#x201D;</p>
</disp-quote>
<p>By contrast, <xref ref-type="fig" rid="fig10">Figure 10</xref> shows the use of REs in topic-shift contexts. Again, these results show a similar trend to those in <xref ref-type="fig" rid="fig8">Figure 8</xref>: monolinguals produce mainly NPs (63.9%), followed by overt pronouns (30.6%). Lower-adv and upper-adv learners show a trend that is rather inverse (though less marked) to monolinguals&#x2019;, by producing overt (58.1, 50%) followed by NPs (41.9, 41.7%), as in (24a, b). Intermediates produce more NPs (54.1%), closely followed by overt (45.9%). Once again, the rather low frequencies in production in all groups may be behind the non-significant differences between each of the learner groups and the monolinguals: non-significant differences (<italic>p</italic>&#x2009;&#x003E;&#x2009;0.05) in most contrasts; marginally non-significant differences (0.05&#x2009;&#x003C;&#x2009;<italic>p</italic>&#x2009;&#x003C;&#x2009;0.10) for NPs in the lower-advanced vs. monolinguals contrast (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;3.23) and the upper-advanced vs. monolinguals contrast (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;2.87); and only one significant difference for overt pronouns in the lower-advanced vs. monolinguals contrast (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;5.13, <italic>p</italic>&#x2009;=&#x2009;0.0234, <italic>h</italic>&#x2009;=&#x2009;0.561).</p>
<list list-type="simple">
<list-item>
<p>(24) a. &#x2026;Ben<sub>i</sub> ten&#x00ED;a memorias de su esposa<sub>j</sub> y su vida con ella<sub>j</sub>. Ella<sub>j</sub> estaba muy bonita&#x2026; [Learner: EN_WR_37_18_5_3_JEP].</p>
</list-item>
</list>
<fig position="float" id="fig10">
<label>Figure 10</label>
<caption>
<p>Production of REs in topic-shift contexts across groups.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g010.tif"/>
</fig>
<disp-quote>
<p>&#x201C;&#x2026;Ben<sub>i</sub> had memories of his wife<sub>j</sub> and his life with her<sub>j</sub>. <bold>She</bold><sub>
<bold>j</bold>
</sub> was very pretty&#x2026;&#x201D;.</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>b. Pilar<sub>i</sub> empieza de desarrollar sus propias opiniones, fuera de su esposo<sub>j</sub>. Su esposo<sub>j</sub> ha empezado una clase donde &#x00D8;<sub>i</sub> aprende&#x2026; [Learner: EN_WR_41_19_5_3_AEM].</p>
</list-item>
</list>
<disp-quote>
<p>&#x201C;Pilar<sub>i</sub> begins to develop her own opinions, outside her husband<sub>j</sub>. <bold>Her husband</bold><sub>
<bold>j</bold>
</sub> has started a class where &#x00D8;<sub>i</sub> learns&#x2026;&#x201D;.</p>
</disp-quote>
<p>The results in <xref ref-type="fig" rid="fig9">Figures 9</xref>, <xref ref-type="fig" rid="fig10">10</xref> thus show that syntactic position (subject/non-subject) overlaps with information status (topic continuity/shift) in such a way that null pronouns typically mark a continuation of topic of the subject antecedent, whereas overt material (NPs and overt pronouns) typically marks a shift in topic. These results empirically demonstrate that the traditional experimental assumption in section 1.2 is correct in corpus production data.</p>
<p>When it comes the syntax-discourse interface, recall that Sorace&#x2019;s IH predicts deficits with AR even at advanced levels. This is confirmed in this study, but only partially since our results show that not all syntax-discourse PAS scenarios are equally problematic. In topic-continuity contexts, despite learners&#x2019; steady increase of null pronouns, the upper-advanced group (76.2%) still significantly differs from monolinguals (95%), but no significant differences were found in topic-shift scenarios with either NPs or overt pronouns. This differential effect is in line with <xref ref-type="bibr" rid="ref26">Lozano&#x2019;s (2016)</xref> Pragmatic Principles Violation Hypothesis (PPHV), originally proposed for general AR in L1 English-L2 Spanish but also confirmed in other scenarios: AR in L1 Greek-L2 Spanish (<xref ref-type="bibr" rid="ref27">Lozano, 2018</xref>; <xref ref-type="bibr" rid="ref32">Margaza and Gavarr&#x00F3;, 2022</xref>); AR in L1 English-L2 Spanish and L1 Spanish-L2 English (<xref ref-type="bibr" rid="ref41">Quesada, 2021</xref>); clitic pronouns in L1 English-L2 Spanish (<xref ref-type="bibr" rid="ref16">Garc&#x00ED;a-Tejada, 2022</xref>); and pragmatic implicatures in L1 Chinese-L2 English (<xref ref-type="bibr" rid="ref10">Feng, 2022</xref>). The PPVH postulates that learners typically obey the pragmatic Principle of Clarity (i.e., they attain native-like knowledge in topic-shift contexts by using full RE forms to avoid ambiguity) but often violate the Principle of Economy (i.e., they produce overt pronouns in topic continuity, which leads to redundancy).</p>
<p>To summarize, the results showed that the choice of REs depends both on (i) the syntactic position of its antecedent (null&#x2794;subject vs. NP/overt&#x2794;non-subject), and (ii) the information status of its antecedent (null&#x2794;topic continuity vs. NP/overt&#x2794;topic shift). <italic>H4</italic> is confirmed as there is a correspondence between syntactic position and information structure in PAS. Finally, the PPVH is confirmed since the most advanced L2ers cannot attain full native-like competence in topic-continuity contexts, but they can in topic-shift contexts.</p>
</sec>
<sec id="sec19">
<label>3.5.</label>
<title><italic>RQ5</italic>/<italic>H5</italic>: cross-linguistic influence</title>
<p>Recall that a null-subject language like Spanish allows null pronominal subjects in all syntactic configurations (coordination and non-coordination), whereas a non-null subject language like English allows them only in topic continuity <italic>and</italic> coordination. If L1 transfer plays a role in PAS, L1 English-L2 Spanish learners are expected to produce null pronouns mostly in contexts where English allows them.</p>
<p>In topic continuity and coordinate syntactic configurations (<italic>cf.</italic> (16b)), all groups produce mostly null pronominal subjects. Learners show a slight increasing trend toward the native norm, though only the upper-advanced group shows native-like knowledge (<xref ref-type="fig" rid="fig11">Figure 11</xref>): intermediates vs. monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;10.54, <italic>p</italic>&#x2009;=&#x2009;0.0012, <italic>h</italic>&#x2009;=&#x2009;1.159); lower-advanced vs. monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;7.78, <italic>p</italic>&#x2009;=&#x2009;0.0053, <italic>h</italic>&#x2009;=&#x2009;0.994); upper-advanced vs. monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;2.93, <italic>p</italic>&#x2009;=&#x2009;0.0870 n.s, <italic>h</italic>&#x2009;=&#x2009;0.613). By contrast, in topic continuity and non-coordinate configurations learners&#x2019; production of null subjects (<italic>cf.</italic> (17) and (18)) is much lower and is always significantly different from monolinguals&#x2019;: intermediates vs. monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;34.27, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.687); lower-advanced vs. monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;23.97, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001, <italic>h</italic>&#x2009;=&#x2009;1.463); upper-advanced vs. monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;5.98, <italic>p</italic>&#x2009;=&#x2009;0.145, <italic>h</italic>&#x2009;=&#x2009;0.715). Additional within-group comparisons<xref ref-type="fn" rid="fn0013"><sup>13</sup></xref> show that Spanish monolinguals&#x2019; production of null pronouns in topic-continuity coordinate vs. non-coordinate configurations is not significantly different, as expected (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;3.38, <italic>p</italic>&#x2009;&#x003E;&#x2009;0.05, n.s), whereas learners&#x2019; production is significantly different: intermediates (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;16.29, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.02); lower-advanced (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;13.29, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.02); upper-advanced (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;5.52, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.02). Results suggest that learners&#x2019; significantly higher use of null pronouns in coordinate than in non-coordinate configurations reflects L1 English influence. Interestingly, learners show a strong gradual trend toward the native norm (intermediate 15.1%, lower-adv 24%, upper-adv 60%), which suggests their sensitivity to the allowability of null pronouns in non-coordinate scenarios increases with proficiency, though their production rates (even at upper-advanced levels) are far from Spanish monolinguals&#x2019;. This confirms learners&#x2019; transfer of null pronouns but an increasing sensitivity to their pragmatics.</p>
<fig position="float" id="fig11">
<label>Figure 11</label>
<caption>
<p>Production of REs in topic continuity and coordinate contexts across groups (only null pronouns plotted).</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g011.tif"/>
</fig>
</sec>
<sec id="sec20">
<label>3.6.</label>
<title><italic>RQ6</italic>: sentential configuration</title>
<p><xref ref-type="table" rid="tab2">Table 2</xref> shows that that the production of intersentential configurations is two thirds (or higher) the production of intrasentential configurations for both L2ers and monolinguals, which indicates that the most natural sentential configuration for AR in PAS scenarios is intersentential, either independent sentences as in <italic>[sentence].[sentence]</italic> or coordinate sentences as in <italic>[sentence]&#x0026;[sentence]</italic>. This clear-cut trend has been rather overlooked in the design of stimuli in previous experimental studies, where intrasentential configurations like <italic>[main [subordinate]]</italic> have been typically the focus of attention. Importantly, only the upper-advanced learners can attain native-like competence as they are not significantly different from Spanish monolinguals (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;0.60, <italic>p</italic>&#x2009;=&#x2009;0.4371, <italic>h</italic>&#x2009;=&#x2009;0.116), whereas the intermediates (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;5.57, <italic>p</italic>&#x2009;=&#x2009;0.0182, <italic>h</italic>&#x2009;=&#x2009;0.338), and the lower-advanced learners (<italic>&#x03C7;</italic><sup>2</sup>&#x2009;=&#x2009;11.31, <italic>p</italic>&#x2009;=&#x2009;0.0008, <italic>h</italic>&#x2009;=&#x2009;0.506) significantly differ from monolinguals.</p>
<table-wrap position="float" id="tab2">
<label>Table 2</label>
<caption>
<p>Syntactic configuration: inter- vs. intra-sentential.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th colspan="2"></th>
<th align="center" valign="top" colspan="2">Intermediate</th>
<th align="center" valign="top" colspan="2">Lower-adv</th>
<th align="center" valign="top" colspan="2">Upper-adv</th>
<th align="center" valign="top" colspan="2">Monolinguals</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" colspan="2">INTER-SENT.</td>
<td align="char" valign="middle" char="." colspan="2">79.3%<break/>(73/92)</td>
<td align="char" valign="middle" char="." colspan="2">85.7%<break/>(72/84)</td>
<td align="char" valign="middle" char="." colspan="2">69.7%<break/>(53/76)</td>
<td align="char" valign="middle" char="." colspan="2">64.3%<break/>(72/112)</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="2">INTRA-SENT.</td>
<td align="left" valign="middle">Main_subord</td>
<td align="char" valign="middle" char="." rowspan="2">20.7% (19/92)</td>
<td align="char" valign="middle" char=".">94.74% (18/19)</td>
<td align="char" valign="middle" char="." rowspan="2">14.3% (12/84)</td>
<td align="char" valign="middle" char=".">83.3%<break/>(10/12)</td>
<td align="char" valign="middle" char="." rowspan="2">30.3% (23/76)</td>
<td align="char" valign="middle" char=".">91.3%<break/>(21/23)</td>
<td align="char" valign="middle" char="." rowspan="2">35.7% (40/112)</td>
<td align="char" valign="middle" char=".">90%<break/>(36/40)</td>
</tr>
<tr>
<td align="left" valign="middle">Subord_main</td>
<td align="char" valign="middle" char=".">5.26% (1/19)</td>
<td align="char" valign="middle" char=".">16.67%<break/>(2/12)</td>
<td align="char" valign="middle" char=".">8.7%<break/>(2/23)</td>
<td align="char" valign="middle" char=".">10%<break/>(4/40)</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Recall that an additional question concerns the order of main and subordinate clauses. <xref ref-type="table" rid="tab2">Table 2</xref> shows the clausal order for the low-frequency intrasentential configurations: Main-subordinate is overwhelmingly more frequent than subordinate-main for both L2ers and monolinguals. Note that no inferential statistics are performed here due to the low frequencies.</p>
<p>In short, in natural production PAS scenarios are overwhelmingly intersentential and, when they happen to be intra-sentential, the most frequent clausal order is main-subordinate. This is so in native and non-native grammars. These findings provide clear tips for those researchers wishing to design experimental PAS configurations that intend to look as natural as possible.</p>
<p>A final consideration is whether the sentential configuration is a factor that modulates the choice of RE in PAS in native Spanish (<xref ref-type="fig" rid="fig12">Figure 12</xref>). Null pronominal subjects are clearly biased toward subject antecedents regardless of the type of sentence (100% intrasentential, 90.5% intersentential), whereas an interesting subdivision of labor is observed when the bias is toward non-subject antecedents: overt pronouns in intrasentential (85.7%) but NPs in intersentential (77.8%), as shown in (25a, b). In other words, topic continuity (subject bias) is marked via null pronouns irrespective of the sentential configuration, but topic shift (non-subject bias) is marked via overt pronouns intrasententially yet via NPs intersententially, which is a finding not reported in the previous literature. Sentential configuration is therefore an additional factor that modulates the division of labor of REs in PAS configurations in native Spanish. This issue merits further investigation in future studies containing larger frequencies of learner and native corpus data.</p>
<list list-type="simple">
<list-item>
<p>(25) a. Intrasentential: overt pronoun biasing toward a non-subject antecedent:</p>
</list-item>
</list>
<fig position="float" id="fig12">
<label>Figure 12</label>
<caption>
<p>Spanish monolinguals&#x2019; production of REs in intrasentential vs. intersentential.</p>
</caption>
<graphic xlink:href="fpsyg-14-1246710-g012.tif"/>
</fig>
<p>Marco<sub>i</sub> est&#x00E1; celoso y &#x00D8;<sub>i</sub> no se adapta bien a esta nueva vida de Ver&#x00F3;nica<sub>j</sub> cuando <bold>ella</bold><sub>j</sub> empieza a tomar a sus amigos como amantes. [Learner: EN_WR_42_21_8_3_LBK].</p>
<disp-quote>
<p>&#x201C;Marco<sub>i</sub> is jealous and &#x00D8;<sub>i</sub> does not adapt well to this new idea of Veronica<sub>j</sub> when <bold>she</bold><sub>
<bold>j</bold>
</sub> starts taking her friends as lovers.&#x201D;</p>
</disp-quote>
<list list-type="simple">
<list-item>
<p>b. Intersentential: NP pronoun biasing toward a non-subject antecedent:</p>
</list-item>
</list>
<p>Ella<sub>i</sub> le<sub>j</sub> tiene mucho cari&#x00F1;o, pero &#x00D8;<sub>i</sub> se niega a desmentir sus votos para estar con &#x00E9;l<sub>j</sub>. <bold>Nacho</bold><sub>j</sub> se deja guiar por un idealismo optimista&#x2026; [Learner: EN_WR_42_21_10_3_LBK].</p>
<disp-quote>
<p>&#x201C;She<sub>i</sub> is very fond of him<sub>j</sub>, but &#x00D8;<sub>i</sub> refuses to deny her vows to be with him<sub>j</sub>. <bold>Nacho</bold><sub>
<bold>j</bold>
</sub> allows himself to be guided by an optimist idealism&#x2026;&#x201D;.</p>
</disp-quote>
</sec>
</sec>
<sec id="sec21">
<label>4.</label>
<title>General discussion and conclusion</title>
<p><italic>RQ1</italic> called into question the PAS as a prototypical way of resolving anaphora in native (and L2) Spanish. Our corpus results confirmed a low production of PAS compared to other AR configurations in natural language production. So, as we found during the corpus sample selection (section 2.2), it is difficult to find PAS in natural narrative production and, in those narrations that include PAS, their frequency is rather low. <xref ref-type="bibr" rid="ref5">Carminati&#x2019;s (2002)</xref> original PAS proposal for Italian has triggered a wealth of experimental studies in many languages and bilingual populations. These studies have blindly tested PAS (and slight variants of it) over and over again but our corpus data show that the PAS is neither a common phenomenon nor prototypical way of resolving anaphora.</p>
<p>Results from <italic>RQ2</italic> confirmed the hypothesis that Spanish native discourse contains mainly null pronominal subjects, while learners&#x2019; production is significantly lower. Importantly, two crucial findings for native Spanish PAS were (i) the rather low production of overt pronouns, which contrasts with their importance in experimental studies, and (ii) the high production of NPs as an anaphoric device, an overlooked factor in experimental studies. Learners&#x2019; PAS behavior ranged from intermediates&#x2019; strong influence from their L1 English (overt pronouns and NPs predominate, with low rates of null pronouns), the indeterminacy of lower advanced learners (production of one third of each RE form), and difficulty to attain native levels by the upper-advanced group since they still produce significantly more overt pronouns than monolinguals, in line with previous corpus research on L2 Spanish dealing with AR in general (<xref ref-type="bibr" rid="ref36">Montrul and Rodr&#x00ED;guez Louro, 2006</xref>; <xref ref-type="bibr" rid="ref25">Lozano, 2009</xref>, <xref ref-type="bibr" rid="ref26">2016</xref>). These findings become more meaningful when we incorporate syntax-discourse factors in PAS, as we will discuss below.</p>
<p><italic>RQ3</italic> addressed a much-debated topic in the literature on native Spanish: the division of labor of RE forms in PAS. Experimental studies report a clear role for null pronouns (they show a strong subject-antecedent bias), yet overt pronouns show a &#x201C;flexible&#x201D; behavior (non-subject- as well as subject-antecedent biases). The corpus data showed a clear division of labor when we consider overtly realized REs together (i.e., overt pronouns and NPs): null pronouns clearly select subject antecedents whereas overt REs clearly select non-subject antecedents. This is quite revealing as NPs were not typically considered in previous experimental PAS studies (except for <xref ref-type="bibr" rid="ref18">Gelormini-Lezama and Almor, 2011</xref>). The relevance of corpus data then becomes clear as a complementary (and needed) source of evidence for experimental data in the study of bilingualism.</p>
<p>As for learners&#x2019; subject antecedents, they start off by showing indeterminacy and no clear PAS strategy in L2 Spanish, but then show a gradual development toward the native norm, but even the upper-advanced group still significantly produces more overt pronouns (and less null pronouns) than monolinguals do to refer to the subject. The results are in line with previous studies regarding development (<xref ref-type="bibr" rid="ref20">Jegerski et al., 2011</xref>) and native-like knowledge but lack of full native-like attainment at advanced levels (<xref ref-type="bibr" rid="ref4">Bel et al., 2016b</xref>; <xref ref-type="bibr" rid="ref8">Clements and Dom&#x00ED;nguez, 2017</xref>). As for learners&#x2019; non-subject antecedents, if we consider overt pronouns and NPs together, the bias is clearer for all groups as overt REs are biased toward non-subject antecedents. So, it seems that the division of labor in learners&#x2019; is clearer from early stages for non-subject antecedents than for subject antecedents. This is not surprising as the antecedent bias is somehow related to the information status (i.e., topic continuity/shift) and topic continuity is more problematic than topic shift, as we discuss next.</p>
<p>Results for <italic>RQ4</italic> confirmed the correspondence between information status and syntactic configuration (i.e., null pronouns&#x2794;subject antecedent/topic continuity; overt pronouns &#x0026; NPs&#x2794;non-subject antecedent/topic shift). Regarding the deficits at the syntax-discourse interface predicted by the IH (<xref ref-type="bibr" rid="ref44">Sorace, 2011</xref>), learners showed deficits, but there were differential effects, as predicted by the PPVH (<xref ref-type="bibr" rid="ref26">Lozano, 2016</xref>): Learners showed native-like behavior in topic-shift, but not in topic-continuity contexts, where even upper-advanced learners redundantly use overt pronouns. In short, learners are more redundant than ambiguous with the PAS.</p>
<p>As for <italic>RQ5</italic>, learners&#x2019; lack of native-like attainment with PAS is also motivated by transfer of null pronominal subjects from their L1 in topic continuity and coordination (and not in topic continuity and non-coordination), a fact also reported by <xref ref-type="bibr" rid="ref34">Mart&#x00ED;n-Villena and Lozano (2020)</xref> for diverse AR contexts. Curiously, the cross-linguistic effect is milder in the opposite direction (L1 Spanish-L2 English), as reported by <xref ref-type="bibr" rid="ref42">Quesada and Lozano (2020)</xref>, so future research could investigate this asymmetry in a more controlled way, e.g., by keeping the task and the type of AR analysis constant but turning the language pairs (L1 English-L2 Spanish vs. L1 Spanish-L2 English) into a variable. Despite transfer, our results also show acquisition effects since learners gradually increase their production of null pronouns in both contexts as their proficiency increases.</p>
<p>As for <italic>RQ6</italic>, our corpus data showed that 1/3 of PAS configurations were intrasentential, of which over 90% were main-subordinate order. Interestingly, some of the studies reviewed above that investigated intrasentential sentences showed contradictory results depending on the order of presentation: main-subordinate order (<xref ref-type="bibr" rid="ref6">Chamorro et al., 2016</xref>; <xref ref-type="bibr" rid="ref3">Bel et al., 2016a</xref>) vs. subordinate-main order (<xref ref-type="bibr" rid="ref12">Filiaci, 2010</xref>; <xref ref-type="bibr" rid="ref13">Filiaci et al., 2014</xref>; <xref ref-type="bibr" rid="ref21">Keating et al., 2016</xref>) (<italic>cf.</italic> <xref ref-type="supplementary-material" rid="SM1">online</xref> <xref ref-type="supplementary-material" rid="SM1">Supplementary material</xref> for exact details). Importantly, our corpus findings also show that null pronouns are clearly biased toward subject antecedents regardless of the type of sentential configuration, but for non-subject antecedents the configuration modulates the choice of RE: overt pronouns are biased toward non-subject antecedents intrasententially whereas NPs do so in intersententially.</p>
<p>The current study presents certain limitations. A larger corpus sample would have probably yielded more stable findings but recall our difficulty in finding texts containing enough PAS examples. Additionally, the tasks certainly lead speakers to narrate different films or describe different famous people, which generates a wide and heterogeneous variety of AR scenarios in the texts produced. This could be minimized by using a prompted task (e.g., the narration of a short Charles Chaplin video clip).</p>
<p>Our findings show the relevance of learner corpus research to investigate theoretically-motivated L2 phenomena (<xref ref-type="bibr" rid="ref29">Lozano, 2021b</xref>). Corpus data have uncovered certain key factors that could be certainly implemented in future experiments. Our research group is currently implementing some of these factors into new experiments (NPs as a form of RE, number of potential antecedents, antecedent-anaphor distance, etc). This is in line with recent claims (<xref ref-type="bibr" rid="ref35">Mendikoetxea and Lozano, 2018</xref>; <xref ref-type="bibr" rid="ref19">Gilquin, 2021</xref>) that the triangulation of <italic>experimental</italic> and <italic>corpus</italic> methods leads to a more well-rounded understanding of complex linguistic phenomena in bilingualism and SLA.</p>
</sec>
<sec sec-type="data-availability" id="sec22">
<title>Data availability statement</title>
<p>Publicly available datasets were analyzed in this study. This data can be found at: <ext-link xlink:href="http://cedel2.learnercorpora.com" ext-link-type="uri">http://cedel2.learnercorpora.com</ext-link>.</p>
</sec>
<sec sec-type="ethics-statement" id="sec23">
<title>Ethics statement</title>
<p>The studies involving humans were approved by Comit&#x00E9; en investigaci&#x00F3;n Humana (Universidad de Granada), 1794/CEIH/2020. The studies were conducted in accordance with the local legislation and institutional requirements. The participants provided their written informed consent to participate in this study.</p>
</sec>
<sec sec-type="author-contributions" id="sec24">
<title>Author contributions</title>
<p>All authors listed have made a substantial, direct, and intellectual contribution to the work and approved it for publication.</p>
</sec>
</body>
<back>
<sec sec-type="funding-information" id="sec25">
<title>Funding</title>
<p>This study was funded by grant number PID2020-113818GB-I00 from MCIN (Ministerio de Ciencia e Innovaci&#x00F3;n), AEI (Agencia Estatal de Investigaci&#x00F3;n) (DOI: <ext-link xlink:href="https://doi.org/10.13039/501100011033" ext-link-type="uri">https://doi.org/10.13039/501100011033</ext-link>) to CL. Publication fees were paid by MCIN/AEI and by the Department of English and German Philology (Universidad de Granada).</p>
</sec>
<sec sec-type="COI-statement" id="sec26">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="sec100" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<sec sec-type="supplementary-material" id="sec27">
<title>Supplementary material</title>
<p>The Supplementary material for this article can be found online at: <ext-link xlink:href="https://www.frontiersin.org/articles/10.3389/fpsyg.2023.1246710/full#supplementary-material" ext-link-type="uri">https://www.frontiersin.org/articles/10.3389/fpsyg.2023.1246710/full#supplementary-material</ext-link></p>
<supplementary-material xlink:href="Data_Sheet_1.docx" id="SM1" mimetype="application/vnd.openxmlformats-officedocument.wordprocessingml.document" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</sec>
<fn-group>
<fn id="fn0001">
<p><sup>1</sup>Also known as PAH (Position of Antecedent Hypothesis).</p>
</fn>
<fn id="fn0002">
<p><sup>2</sup>Note that in all the experimental studies under review, the stimuli always contain two potential antecedents (one in subject position, another in non-subject position). The advantage of using corpus data is that in natural production PAS structures typically contain more antecedents and in different syntactic positions (see sections 2.4 and 3.1).</p>
</fn>
<fn id="fn0003">
<p><sup>3</sup>The authors compared discourse-coordinating (<italic>mientras</italic> &#x201C;while&#x201D;) vs. -subordinating (<italic>cuando</italic> &#x201C;when&#x201D;/<italic>despu&#x00E9;s de que</italic> &#x201C;after&#x201D;/<italic>desde que</italic> &#x2018;since&#x2019;) conjuntions. For brevity, we discuss the discourse-coordination results only.</p>
</fn>
<fn id="fn0004">
<p><sup>4</sup>RT differences&#x2009;= <italic>subject antecedent</italic> (RT of object region + RT of PP region) &#x2013; <italic>object antecedent</italic> (RT of object region + RT of PP region).</p>
</fn>
<fn id="fn0005">
<p><sup>5</sup>RT differences&#x2009;= <italic>subject antecedent</italic> (RT of main clause) &#x2013; <italic>object antecedent</italic> (RT of main clause).</p>
</fn>
<fn id="fn0006">
<p><sup>6</sup>The discursive/pragmatic properties of AR are not fully acquired until around 15&#x2009;years of age (<xref ref-type="bibr" rid="ref43">Shin and Smith Cairns, 2012</xref>), so evidence from these teenage monolinguals should be taken cautiously.</p>
</fn>
<fn id="fn0007">
<p><sup>7</sup>We incorporate NPs and repeated proper Ns as type of anaphoric form, hence we use the wider term Referential Expressions (REs) to include all forms (overt/null pronouns, NPs, repeated Ns), instead of the more restrictive term anaphoric forms.</p>
</fn>
<fn id="fn0008">
<p><sup>8</sup><ext-link xlink:href="http://learnercorpora.com" ext-link-type="uri">http://learnercorpora.com</ext-link>
</p>
</fn>
<fn id="fn0009">
<p><sup>9</sup>Only monolinguals from Spain were chosen since in certain varieties (Mexican, Caribbean, Puerto Rican), overt pronouns mark topic continuity (<xref ref-type="bibr" rid="ref14">Flores-Ferr&#x00E1;n, 2004</xref>).</p>
</fn>
<fn id="fn0010">
<p><sup>10</sup><ext-link xlink:href="http://www.corpustool.com" ext-link-type="uri">http://www.corpustool.com</ext-link>
</p>
</fn>
<fn id="fn0011">
<p><sup>11</sup>The original tagging scheme included a richer tagset with more tags that are not analyzed in this study due to space limitations &#x2013;see <xref ref-type="bibr" rid="ref41">Quesada (2021)</xref> for details.</p>
</fn>
<fn id="fn0012">
<p><sup>12</sup>After each corpus example, we provide in square brackets the filename from the CEDEL2 corpus (<ext-link xlink:href="http://cedel2.learnercorpora.com" ext-link-type="uri">http://cedel2.learnercorpora.com</ext-link>).</p>
</fn>
<fn id="fn0013">
<p><sup>13</sup>The latest release of UAM Corpus Tool (version 6.2j, February 2023) does not allow complex within-group comparisons, so we used an earlier release (version 3.3x, August 2021) to do the analysis, though note that version 3.3x reports <italic>p</italic> value ranges (non significant <italic>p</italic> &#x003E;&#x2009;0.05, significant <italic>p</italic> &#x003C;&#x2009;0.05, highly significant <italic>p</italic> &#x003C;&#x2009;0.02) and does not report effect size.</p>
</fn>
</fn-group>
<ref-list>
<title>References</title>
<ref id="ref1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Alonso-Ovalle</surname> <given-names>L.</given-names></name> <name><surname>Fern&#x00E1;ndez-Solera</surname> <given-names>S.</given-names></name> <name><surname>Frazier</surname> <given-names>L.</given-names></name> <name><surname>Clifton</surname> <given-names>C.</given-names></name></person-group> (<year>2002</year>). <article-title>Null vs. overt pronouns and the topic-focus articulation in Spanish: 2704</article-title>. <source>Ital. J. Linguis.</source> <volume>14</volume>, <fpage>151</fpage>&#x2013;<lpage>170</lpage>.</citation></ref>
<ref id="ref2"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Bel</surname> <given-names>A.</given-names></name> <name><surname>Garc&#x00ED;a-Alcaraz</surname> <given-names>E.</given-names></name></person-group> (<year>2015</year>). &#x201C;<article-title>Subject pronouns in the L2 Spanish of Moroccan Arabic speakers: evidence from bilingual and second language learners</article-title>&#x201D; in <source>The Acquisition of Spanish in understudied language pairings</source>. eds. <person-group person-group-type="editor"><name><surname>Judy</surname> <given-names>T.</given-names></name> <name><surname>Perpi&#x00F1;&#x00E1;n</surname> <given-names>S.</given-names></name></person-group> (<publisher-loc>Amsterdam</publisher-loc>, <publisher-loc>John Benjamins</publisher-loc>)</citation></ref>
<ref id="ref3"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Bel</surname> <given-names>A.</given-names></name> <name><surname>Garc&#x00ED;a-Alcaraz</surname> <given-names>E.</given-names></name> <name><surname>Rosado</surname> <given-names>E.</given-names></name></person-group> (<year>2016a</year>). <article-title>Reference comprehension and production in bilingual Spanish: the view from null subject languages</article-title>. In <person-group person-group-type="editor"><name><surname>Fuente</surname> <given-names>A. A.</given-names><prefix>de la</prefix></name> <name><surname>Valenzuela</surname> <given-names>E.</given-names></name> <name><surname>Mart&#x00ED;nez Sanz</surname> <given-names>C.</given-names></name></person-group> (Eds.), <source>Language acquisition beyond parameters</source> <publisher-loc>Amsterdam</publisher-loc>: <publisher-loc>John Benjamins</publisher-loc>.</citation></ref>
<ref id="ref4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bel</surname> <given-names>A.</given-names></name> <name><surname>Sagarra</surname> <given-names>N.</given-names></name> <name><surname>Com&#x00ED;nguez</surname> <given-names>J. P.</given-names></name> <name><surname>Garc&#x00ED;a-Alcaraz</surname> <given-names>E.</given-names></name></person-group> (<year>2016b</year>). <article-title>Transfer and proficiency effects in L2 processing of subject anaphora</article-title>. <source>Lingua</source> <volume>184</volume>, <fpage>134</fpage>&#x2013;<lpage>159</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2016.07.001</pub-id></citation></ref>
<ref id="ref5"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Carminati</surname> <given-names>M. N.</given-names></name></person-group> (<year>2002</year>). <source>The processing of Italian subject pronouns</source> <publisher-loc>Boston</publisher-loc>: <publisher-name>University of Massachusetts</publisher-name>.</citation></ref>
<ref id="ref6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chamorro</surname> <given-names>G.</given-names></name> <name><surname>Sorace</surname> <given-names>A.</given-names></name> <name><surname>Sturt</surname> <given-names>P.</given-names></name></person-group> (<year>2016</year>). <article-title>What is the source of L1 attrition? The effect of recent L1 re-exposure on Spanish speakers under L1 attrition</article-title>. <source>Biling. Lang. Congn.</source> <volume>19</volume>, <fpage>520</fpage>&#x2013;<lpage>532</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S1366728915000152</pub-id></citation></ref>
<ref id="ref7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clahsen</surname> <given-names>H.</given-names></name> <name><surname>Felser</surname> <given-names>C.</given-names></name></person-group> (<year>2006</year>). <article-title>How native-like is non-native language processing?</article-title> <source>Trends Cogn. Sci.</source> <volume>10</volume>, <fpage>564</fpage>&#x2013;<lpage>570</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.tics.2006.10.002</pub-id>, PMID: <pub-id pub-id-type="pmid">17071131</pub-id></citation></ref>
<ref id="ref8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Clements</surname> <given-names>M.</given-names></name> <name><surname>Dom&#x00ED;nguez</surname> <given-names>L.</given-names></name></person-group> (<year>2017</year>). <article-title>Reexamining the acquisition of null subject pronouns in a second language: focus on referential and pragmatic constraints</article-title>. <source>Linguist. Approach. Biling.</source> <volume>7</volume>, <fpage>33</fpage>&#x2013;<lpage>62</lpage>. doi: <pub-id pub-id-type="doi">10.1075/lab.14012.cle</pub-id></citation></ref>
<ref id="ref9"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Egbert</surname> <given-names>J.</given-names></name> <name><surname>Larsson</surname> <given-names>T.</given-names></name> <name><surname>Biber</surname> <given-names>D.</given-names></name></person-group> (<year>2020</year>). <source>Doing linguistics with a corpus: Methodological considerations for the everyday user</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="ref10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Feng</surname> <given-names>S.</given-names></name></person-group> (<year>2022</year>). <article-title>L2 tolerance of pragmatic violations of informativeness: evidence from ad hoc implicatures and contrastive inference</article-title>. <source>Linguist. Appr. Bil.</source> doi: <pub-id pub-id-type="doi">10.1075/lab.21064.fen</pub-id></citation></ref>
<ref id="ref11"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Fern&#x00E1;ndez</surname> <given-names>E. M.</given-names></name> <name><surname>Smith Cairns</surname> <given-names>H.</given-names></name></person-group> (<year>2011</year>). <source>Fundamentals of psycholinguistics</source>. <publisher-loc>Hoboken</publisher-loc>: <publisher-name>Wiley-Blackwell</publisher-name>.</citation></ref>
<ref id="ref12"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Filiaci</surname> <given-names>F.</given-names></name></person-group> (<year>2010</year>). &#x201C;<article-title>Null and overt subject biases in Spanish and Italian: a cross-linguistic comparison</article-title>&#x201D; in <source>Selected proceedings of the 12th Hispanic linguistics symposium</source>. eds. <person-group person-group-type="editor"><name><surname>Borgonovo</surname> <given-names>C.</given-names></name> <name><surname>Espa&#x00F1;ol-Echevarr&#x00ED;a</surname> <given-names>M.</given-names></name> <name><surname>Pr&#x00E9;vost</surname> <given-names>P.</given-names></name></person-group> (<publisher-loc>Hoboken</publisher-loc>: <publisher-name>Cascadilla Press</publisher-name>)</citation></ref>
<ref id="ref13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Filiaci</surname> <given-names>F.</given-names></name> <name><surname>Sorace</surname> <given-names>A.</given-names></name> <name><surname>Carreiras</surname> <given-names>M.</given-names></name></person-group> (<year>2014</year>). <article-title>Anaphoric biases of null and overt subjects in Italian and Spanish: a cross-linguistic comparison</article-title>. <source>Lang. Cogn. Neurosci.</source> <volume>29</volume>, <fpage>825</fpage>&#x2013;<lpage>843</lpage>. doi: <pub-id pub-id-type="doi">10.1080/01690965.2013.801502</pub-id></citation></ref>
<ref id="ref14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Flores-Ferr&#x00E1;n</surname> <given-names>N.</given-names></name></person-group> (<year>2004</year>). <article-title>Spanish subject personal pronoun use in New York City Puerto Ricans: can we rest the case of English contact?</article-title> <source>Lang. Var. Chang.</source> <volume>16</volume>, <fpage>49</fpage>&#x2013;<lpage>73</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0954394504161048</pub-id></citation></ref>
<ref id="ref15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Garc&#x00ED;a-Alcaraz</surname> <given-names>E.</given-names></name> <name><surname>Bel</surname> <given-names>A.</given-names></name></person-group> (<year>2019</year>). <article-title>Does empirical data from bilingual and native Spanish corpora meet linguistic theory? The role of discourse context in variation of subject expression</article-title>. <source>Appl. Linguist. Rev.</source> <volume>10</volume>, <fpage>491</fpage>&#x2013;<lpage>515</lpage>. doi: <pub-id pub-id-type="doi">10.1515/applirev-2017-0101</pub-id></citation></ref>
<ref id="ref16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Garc&#x00ED;a-Tejada</surname> <given-names>A.</given-names></name></person-group> (<year>2022</year>). <article-title>Direct object anaphora resolution in L1 English-L2 Spanish: referring clitics and DPs</article-title>. <source>Revista Espa&#x00F1;ola de Ling&#x00FC;&#x00ED;stica Aplicada</source>. doi: <pub-id pub-id-type="doi">10.1075/resla.22014.gar</pub-id></citation></ref>
<ref id="ref17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Geber</surname> <given-names>D.</given-names></name></person-group> (<year>2006</year>). <article-title>Processing subject pronouns in relation to non-canonical (quirky) constructions</article-title>. <source>Ottawa Pap. Linguist.</source> <volume>34</volume>, <fpage>47</fpage>&#x2013;<lpage>61</lpage>.</citation></ref>
<ref id="ref18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gelormini-Lezama</surname> <given-names>C.</given-names></name> <name><surname>Almor</surname> <given-names>A.</given-names></name></person-group> (<year>2011</year>). <article-title>Repeated names, overt pronouns, and null pronouns in Spanish</article-title>. <source>Lang. Cogn. Process.</source> <volume>26</volume>, <fpage>437</fpage>&#x2013;<lpage>454</lpage>. doi: <pub-id pub-id-type="doi">10.1080/01690965.2010.495234</pub-id>, PMID: <pub-id pub-id-type="pmid">21552376</pub-id></citation></ref>
<ref id="ref19"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Gilquin</surname> <given-names>G.</given-names></name></person-group> (<year>2021</year>). &#x201C;<article-title>Combining learner corpora and experimental methods</article-title>&#x201D; in <source>The Routledge handbook of second language acquisition and corpora</source>. eds. <person-group person-group-type="editor"><name><surname>Tracy-Ventura</surname> <given-names>N.</given-names></name> <name><surname>Paquot</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>)</citation></ref>
<ref id="ref20"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jegerski</surname> <given-names>J.</given-names></name> <name><surname>Van Patten</surname> <given-names>B.</given-names></name> <name><surname>Keating</surname> <given-names>G. D.</given-names></name></person-group> (<year>2011</year>). <article-title>Cross-linguistic variation and the acquisition of pronominal reference in L2 Spanish</article-title>. <source>Second. Lang. Res.</source> <volume>27</volume>, <fpage>481</fpage>&#x2013;<lpage>507</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0267658311406033</pub-id></citation></ref>
<ref id="ref21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Keating</surname> <given-names>G. D.</given-names></name> <name><surname>Jegerski</surname> <given-names>J.</given-names></name> <name><surname>Vanpatten</surname> <given-names>B.</given-names></name></person-group> (<year>2016</year>). <article-title>Online processing of subject pronouns in monolingual and heritage bilingual speakers of Mexican Spanish</article-title>. <source>Biling. Lang. Congn.</source> <volume>19</volume>, <fpage>36</fpage>&#x2013;<lpage>49</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S1366728914000418</pub-id></citation></ref>
<ref id="ref22"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Keating</surname> <given-names>G. D.</given-names></name> <name><surname>Vanpatten</surname> <given-names>B.</given-names></name> <name><surname>Jegerski</surname> <given-names>J.</given-names></name></person-group> (<year>2011</year>). <article-title>Who was walking on the beach?: anaphora resolution in Spanish heritage speakers and adult second language learners</article-title>. <source>Stud. Second. Lang. Acquis.</source> <volume>33</volume>, <fpage>193</fpage>&#x2013;<lpage>221</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0272263110000732</pub-id></citation></ref>
<ref id="ref23"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Kra&#x0161;</surname> <given-names>T.</given-names></name></person-group> (<year>2008a</year>). <article-title>Anaphora resolution in Croatian: psycholinguistic evidence from native speakers</article-title>. In <person-group person-group-type="editor"><name><surname>Tadi&#x0107;</surname> <given-names>M.</given-names></name> <name><surname>Dimitrova-Vulchanova</surname> <given-names>M.</given-names></name> <name><surname>Koeva</surname> <given-names>S.</given-names></name></person-group> (Eds.), <conf-name>Proceedings of the Sixth International Conference &#x201C;Formal Approaches to South Slavic and Balkan languages</conf-name> <conf-loc>Zagreb, Croatia: Croatian Language Technologies Society &#x2013; Faculty of Humanities and Social Sciences</conf-loc>.</citation></ref>
<ref id="ref24"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Kra&#x0161;</surname> <given-names>T.</given-names></name></person-group> (<year>2008b</year>). &#x201C;<article-title>Anaphora resolution in near-native Italian grammars: evidence from native speakers of Croatian</article-title>&#x201D; in <source>EUROSLA yearbook <italic>8</italic></source>. eds. <person-group person-group-type="editor"><name><surname>Roberts</surname> <given-names>L.</given-names></name> <name><surname>Myles</surname> <given-names>F.</given-names></name> <name><surname>David</surname> <given-names>A.</given-names></name></person-group> (<publisher-loc>Hoboken</publisher-loc>: <publisher-name>John Benjamins</publisher-name>)</citation></ref>
<ref id="ref25"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2009</year>). &#x201C;<article-title>Selective deficits at the syntax-discourse interface: evidence from the CEDEL2 corpus</article-title>&#x201D; in <source>Representational deficits in SLA: Studies in honor of Roger Hawkins</source>. eds. <person-group person-group-type="editor"><name><surname>Snape</surname> <given-names>N.</given-names></name> <name><surname>Leung</surname> <given-names>Y. I.</given-names></name> <name><surname>Sharwood Smith</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Hoboken</publisher-loc>: <publisher-name>John Benjamins</publisher-name>)</citation></ref>
<ref id="ref26"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2016</year>). &#x201C;<article-title>Pragmatic principles in anaphora resolution at the syntax-discourse interface: advanced English learners of Spanish in the CEDEL2 corpus</article-title>&#x201D; in <source>Spanish learner corpus research: State of the art and perspectives</source>. ed. <person-group person-group-type="editor"><name><surname>Alonso-Ramos</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Hoboken</publisher-loc>: <publisher-name>John Benjamins</publisher-name>)</citation></ref>
<ref id="ref27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2018</year>). <article-title>The development of anaphora resolution at the syntax-discourse Interface: pronominal subjects in Greek learners of Spanish</article-title>. <source>J. Psycholinguist. Res.</source> <volume>47</volume>, <fpage>411</fpage>&#x2013;<lpage>430</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10936-017-9541-8</pub-id>, PMID: <pub-id pub-id-type="pmid">29197978</pub-id></citation></ref>
<ref id="ref28"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2021a</year>). &#x201C;<article-title>Anaphora resolution in second language acquisition</article-title>&#x201D; in <source>Oxford bibliographies in linguistics</source>. ed. <person-group person-group-type="editor"><name><surname>Aronoff</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>)</citation></ref>
<ref id="ref29"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2021b</year>). &#x201C;<article-title>Generative approaches</article-title>&#x201D; in <source>The Routledge handbook of second language acquisition and corpora</source>. eds. <person-group person-group-type="editor"><name><surname>Tracy-Ventura</surname> <given-names>N.</given-names></name> <name><surname>Paquot</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>)</citation></ref>
<ref id="ref30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2022</year>). <article-title>CEDEL2: design, compilation and web interface of an online corpus for L2 Spanish acquisition research</article-title>. <source>Second. Lang. Res.</source> <volume>38</volume>, <fpage>965</fpage>&#x2013;<lpage>983</lpage>. doi: <pub-id pub-id-type="doi">10.1177/02676583211050522</pub-id></citation></ref>
<ref id="ref31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mac Donald</surname> <given-names>M.</given-names></name></person-group> (<year>2013</year>). <article-title>How language production shapes language form and comprehension</article-title>. <source>Front. Psychol.</source> <volume>4</volume>:<fpage>226</fpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2013.00226</pub-id>, PMID: <pub-id pub-id-type="pmid">23637689</pub-id></citation></ref>
<ref id="ref32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Margaza</surname> <given-names>P.</given-names></name> <name><surname>Gavarr&#x00F3;</surname> <given-names>A.</given-names></name></person-group> (<year>2022</year>). <article-title>The distribution of subjects in L2 Spanish by Greek learners</article-title>. <source>Front. Psychol.</source> <volume>12</volume>:<fpage>794587</fpage>. doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2021.794587</pub-id>, PMID: <pub-id pub-id-type="pmid">35126245</pub-id></citation></ref>
<ref id="ref33"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Mart&#x00ED;n-Villena</surname> <given-names>F.</given-names></name></person-group> (<year>2023</year>). <article-title>L1 morphosyntactic attrition at the early stages: Evidence from production, interpretation, and processing of subject referring expressions in L1 Spanish-L2 English instructed and immersed bilinguals</article-title> [PhD dissertation, Universidad de Granada]. Available at: <ext-link xlink:href="https://hdl.handle.net/10481/81920" ext-link-type="uri">https://hdl.handle.net/10481/81920</ext-link></citation></ref>
<ref id="ref34"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Mart&#x00ED;n-Villena</surname> <given-names>F.</given-names></name> <name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). &#x201C;<article-title>Anaphora resolution in topic continuity: evidence from L1 English&#x2013;L2 Spanish data in the CEDEL2 corpus</article-title>&#x201D; in <source>Referring in a second language: Studies on reference to person in a multilingual world</source>. eds. <person-group person-group-type="editor"><name><surname>Ryan</surname> <given-names>J.</given-names></name> <name><surname>Crosthwaite</surname> <given-names>P.</given-names></name></person-group> (<publisher-loc>London</publisher-loc>: <publisher-name>Routledge</publisher-name>)</citation></ref>
<ref id="ref35"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Mendikoetxea</surname> <given-names>A.</given-names></name> <name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2018</year>). <article-title>From corpora to experiments: methodological triangulation in the study of word order at the interfaces in adult late bilinguals (L2 learners)</article-title>. <source>J. Psycholinguist. Res.</source> <volume>47</volume>, <fpage>871</fpage>&#x2013;<lpage>898</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10936-018-9560-0</pub-id>, PMID: <pub-id pub-id-type="pmid">29404914</pub-id></citation></ref>
<ref id="ref36"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Montrul</surname> <given-names>S.</given-names></name> <name><surname>Rodr&#x00ED;guez Louro</surname> <given-names>C.</given-names></name></person-group> (<year>2006</year>). &#x201C;<article-title>Beyond the syntax of the null subject parameter: a look at the discourse-pragmatic distribution of null and overt subjects by L2 learners of Spanish</article-title>&#x201D; in <source>Language acquisition and language disorders</source>. eds. <person-group person-group-type="editor"><name><surname>Torrens</surname> <given-names>V.</given-names></name> <name><surname>Escobar</surname> <given-names>L.</given-names></name></person-group> (<publisher-loc>Hoboken</publisher-loc>: <publisher-name>John Benjamins</publisher-name>)</citation></ref>
<ref id="ref37"><citation citation-type="book"><person-group person-group-type="author"><name><surname>O&#x2019;Donnell</surname> <given-names>M.</given-names></name></person-group> (<year>2009</year>). &#x201C;<article-title>The UAM corpus tool: software for corpus annotation and exploration</article-title>&#x201D; in <source>Applied linguistics now: Understanding language and mind/La Ling&#x00FC;&#x00ED;stica Aplicada actual: Comprendiendo el Lenguaje y la Mente</source>. eds. <person-group person-group-type="editor"><name><surname>Bretones</surname> <given-names>C. M.</given-names></name> <etal/></person-group>. (<publisher-loc>Almer&#x00ED;a</publisher-loc>: <publisher-name>Universidad de Almer&#x00ED;a</publisher-name>)</citation></ref>
<ref id="ref38"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Papadopoulou</surname> <given-names>D.</given-names></name> <name><surname>Peristeri</surname> <given-names>E.</given-names></name> <name><surname>Plemenou</surname> <given-names>E.</given-names></name> <name><surname>Marinis</surname> <given-names>T.</given-names></name> <name><surname>Tsimpli</surname> <given-names>I.</given-names></name></person-group> (<year>2015</year>). <article-title>Pronoun ambiguity resolution in Greek: evidence from monolingual adults and children</article-title>. <source>Lingua</source> <volume>155</volume>, <fpage>98</fpage>&#x2013;<lpage>120</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2014.09.006</pub-id></citation></ref>
<ref id="ref39"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Pickering</surname> <given-names>M. J.</given-names></name> <name><surname>van Gompel</surname> <given-names>P. G.</given-names></name></person-group> (<year>2006</year>). &#x201C;<article-title>Syntactic parsing</article-title>&#x201D; in <source>Handbook of psycholinguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Traxler</surname> <given-names>M.</given-names></name> <name><surname>Gernsbacher</surname> <given-names>M. A.</given-names></name></person-group>. <edition>2nd</edition> ed (<publisher-loc>Cambridge</publisher-loc>: <publisher-name>Academic Press</publisher-name>)</citation></ref>
<ref id="ref40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Prentza</surname> <given-names>A.</given-names></name> <name><surname>Tsimpli</surname> <given-names>I.-M.</given-names></name></person-group> (<year>2013</year>). <article-title>Resolution of pronominal ambiguity in Greek: syntax and pragmatics</article-title>. <source>Stud. Greek Linguist.</source> <volume>33</volume>, <fpage>197</fpage>&#x2013;<lpage>208</lpage>.</citation></ref>
<ref id="ref41"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Quesada</surname> <given-names>T.</given-names></name></person-group> (<year>2021</year>). <article-title>Studies on anaphora resolution in L1 Spanish-L2 English and L1 English-L2 Spanish adult learners: Combining corpus and experimental methods</article-title> [PhD dissertation, Universidad de Granada]. Available at: <ext-link xlink:href="http://hdl.handle.net/10481/72052" ext-link-type="uri">http://hdl.handle.net/10481/72052</ext-link></citation></ref>
<ref id="ref42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Quesada</surname> <given-names>T.</given-names></name> <name><surname>Lozano</surname> <given-names>C.</given-names></name></person-group> (<year>2020</year>). <article-title>Which factors determine the choice of referential expressions in L2 English discourse? A multifactorial study from the COREFL corpus</article-title>. <source>Stud. Second. Lang. Acquis.</source> <volume>42</volume>, <fpage>959</fpage>&#x2013;<lpage>986</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0272263120000224</pub-id></citation></ref>
<ref id="ref43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shin</surname> <given-names>N. L.</given-names></name> <name><surname>Smith Cairns</surname> <given-names>H.</given-names></name></person-group> (<year>2012</year>). <article-title>The development of NP selection in school-age children: reference and Spanish subject pronouns</article-title>. <source>Lang. Acquis.</source> <volume>19</volume>, <fpage>3</fpage>&#x2013;<lpage>38</lpage>. doi: <pub-id pub-id-type="doi">10.1080/10489223.2012.633846</pub-id></citation></ref>
<ref id="ref44"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Sorace</surname> <given-names>A.</given-names></name></person-group> (<year>2011</year>). <article-title>Pinning down the concept of &#x201C;interface&#x201D; in bilingualism</article-title>. <source>Linguist. Approach. Bilingual.</source> <volume>1</volume>, <fpage>1</fpage>&#x2013;<lpage>33</lpage>. doi: <pub-id pub-id-type="doi">10.1075/lab.1.1.01sor</pub-id></citation></ref>
<ref id="ref45"><citation citation-type="book"><person-group person-group-type="author"><collab id="coll1">University of Wisconsin</collab></person-group>. (<year>1998</year>). <source>The University of Wisconsin College-Level Placement Test: Spanish (grammar) form 96M</source>. <publisher-loc>Madison</publisher-loc>: <publisher-name>University of Wisconsin Press</publisher-name>.</citation></ref>
</ref-list>
</back>
</article>