<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2024.1384629</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Language transfer in L2 academic writings: a dependency grammar approach</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name><surname>Bi</surname> <given-names>Yude</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<role content-type="https://credit.niso.org/contributor-roles/conceptualization/"/>
<role content-type="https://credit.niso.org/contributor-roles/supervision/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-review-editing/"/>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name><surname>Tan</surname> <given-names>Hua</given-names></name>
<xref ref-type="aff" rid="aff1"><sup>1</sup></xref>
<xref ref-type="aff" rid="aff2"><sup>2</sup></xref>
<xref ref-type="corresp" rid="c001"><sup>&#x002A;</sup></xref>
<xref rid="fn0014" ref-type="author-notes"><sup>&#x2020;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/2199882/overview"/>
<role content-type="https://credit.niso.org/contributor-roles/data-curation/"/>
<role content-type="https://credit.niso.org/contributor-roles/resources/"/>
<role content-type="https://credit.niso.org/contributor-roles/software/"/>
<role content-type="https://credit.niso.org/contributor-roles/visualization/"/>
<role content-type="https://credit.niso.org/contributor-roles/writing-original-draft/"/>
</contrib>
</contrib-group>
<aff id="aff1"><sup>1</sup><institution>Fudan University</institution>, <addr-line>Shanghai</addr-line>, <country>China</country></aff>
<aff id="aff2"><sup>2</sup><institution>Central China Normal University</institution>, <addr-line>Wuhan</addr-line>, <country>China</country></aff>
<author-notes>
<fn fn-type="edited-by" id="fn0001">
<p>Edited by: Julie Franck, University of Geneva, Switzerland</p>
</fn>
<fn fn-type="edited-by" id="fn0002">
<p>Reviewed by: Anastasia Paspali, Aristotle University of Thessaloniki, Greece</p>
<p>Hang Su, Sichuan International Studies University, China</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: Hua Tan, <email>jacktanhua@ccnu.edu.cn</email></corresp>
<fn fn-type="present-address" id="fn0014">
<p><sup>&#x2020;</sup>Present address: Hua Tan,Key Laboratory of Language Science and Multilingual Artificial Intelligence, Shanghai International Studies University, Shanghai, China</p>
</fn>
</author-notes>
<pub-date pub-type="epub">
<day>09</day>
<month>05</month>
<year>2024</year>
</pub-date>
<pub-date pub-type="collection">
<year>2024</year>
</pub-date>
<volume>15</volume>
<elocation-id>1384629</elocation-id>
<history>
<date date-type="received">
<day>10</day>
<month>02</month>
<year>2024</year>
</date>
<date date-type="accepted">
<day>25</day>
<month>04</month>
<year>2024</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2024 Bi and Tan.</copyright-statement>
<copyright-year>2024</copyright-year>
<copyright-holder>Bi and Tan</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>Dependency distance (DD) is an important factor in language processing and can affect the ease with which a sentence is understood. Previous studies have investigated the role of DD in L2 writing, but little is known about how the native language influences DD in L2 academic writing. This study is probably the first one that investigates, though a large dataset of over 400 million words, whether the native language of L2 writers influences the DD in their academic writings. Using a dataset of over 2.2 million abstracts of articles downloaded from Scopus in the fields of Arts &#x0026; Humanities and Social Sciences, the study analyzes the DD patterns, parsed by the latest version of the syntactic parser Stanford Corenlp 4.5.5, in the academic writing of L2 learners from different language backgrounds. It is found that native languages influence the DD of English L2 academic writings. When the mean dependency distance (MDD) of native languages is much longer than that of native English, the MDD of their English L2 academic writings will be much longer than that of English native academic writings. The findings of this study will deepen our insights into the influence of native language transfer on L2 academic writing, potentially shaping pedagogical strategies in L2 academic writing education.</p>
</abstract>
<kwd-group>
<kwd>dependency distance</kwd>
<kwd>English L2 academic writing</kwd>
<kwd>native language transfer</kwd>
<kwd>corpus analysis</kwd>
<kwd>dependency direction</kwd>
</kwd-group>
<counts>
<fig-count count="7"/>
<table-count count="11"/>
<equation-count count="7"/>
<ref-count count="53"/>
<page-count count="14"/>
<word-count count="9525"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Psychology of Language</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec sec-type="intro" id="sec1">
<label>1</label>
<title>Introduction</title>
<p>Academic writing in a second language (L2) poses many challenges for L2 learners. One important aspect of academic writing quality is syntactic complexity (<xref ref-type="bibr" rid="ref34">Lu, 2011</xref>), which can be measured by dependency distance (DD)&#x2014;the linear distance between syntactically related words in a sentence (<xref ref-type="bibr" rid="ref9001">Tesni&#x00E8;re, 1959</xref>; <xref ref-type="bibr" rid="ref28">Liu, 2007</xref>, <xref ref-type="bibr" rid="ref29">2008</xref>; <xref ref-type="bibr" rid="ref20">Hudson, 2010</xref>; <xref ref-type="bibr" rid="ref32">Liu et al., 2017</xref>, <xref ref-type="bibr" rid="ref33">2022</xref>). In their exploration of DD, researchers have also examined dependency direction (<xref ref-type="bibr" rid="ref31">Liu et al., 2009</xref>; <xref ref-type="bibr" rid="ref30">Liu, 2010</xref>; <xref ref-type="bibr" rid="ref21">Jiang and Liu, 2015</xref>; <xref ref-type="bibr" rid="ref49">Wang and Liu, 2017</xref>; <xref ref-type="bibr" rid="ref9">Fan and Jiang, 2019</xref>)&#x2014;a concept that delineates the positional relationship between a governor and its dependent within syntactically connected word pairs, specifically whether the governor appears after or before its dependent. DD has been recognized as a valid measure of syntactic complexity and language comprehension difficulty (<xref ref-type="bibr" rid="ref29">Liu, 2008</xref>; <xref ref-type="bibr" rid="ref41">Oya, 2011</xref>). Research has found that writers tend to minimize DD in the writings with their native languages (<xref ref-type="bibr" rid="ref45">Temperley, 2007</xref>, <xref ref-type="bibr" rid="ref46">2008</xref>; <xref ref-type="bibr" rid="ref11">Futrell et al., 2015</xref>; <xref ref-type="bibr" rid="ref47">Temperley and Gildea, 2018</xref>; <xref ref-type="bibr" rid="ref26">Lei and Wen, 2019</xref>; <xref ref-type="bibr" rid="ref36">Lu and Liu, 2020</xref>), resulting in more locally coherent sentences. However, less is known about the DDs of English L2 academic writings and whether the writers&#x2019; native languages influence the DD of their L2 academic writings.</p>
<p>This study intends to investigate whether the DDs of English L2 academic writing are affected by the writers&#x2019; native languages. Specifically, it compares the DDs of the abstracts of journal articles written by English L2 users and English native speakers. English is found to be different from other languages, like French, Spanish, Korean, and Arabic, in thought patterns and rhetorical structures (<xref ref-type="bibr" rid="ref23">Kaplan, 1966</xref>), which may impact the DDs in L2 writing. However, L2 writing may also be shaped by universal pressures for efficient processing, driving DDs toward a common optimal range (<xref ref-type="bibr" rid="ref32">Liu et al., 2017</xref>) (See section 2 for further explanation).</p>
<p>To test these accounts, we analyzed the DDs of English academic writings by English L2 users and native speakers. We extracted the DDs from each text using syntactic parsing and compared the distributions statistically. This allows us to determine if native language background influences DD in English L2 academic writing, shedding light on how linguistic backgrounds impact L2 syntactic structures in academic writing. Such insights could contribute significantly to our understanding of language acquisition and the challenges faced by individuals writing in an L2 academic context. The findings will have implications for understanding the role of native language transfer in English L2 writing.</p>
<sec id="sec2">
<label>1.1</label>
<title>Previous research</title>
<p>Syntactic complexity, which involves the range and sophistication of syntactic structures, is considered a key dimension of academic writing development and quality (<xref ref-type="bibr" rid="ref39">Ortega, 2003</xref>; <xref ref-type="bibr" rid="ref34">Lu, 2011</xref>). A quantitative metric that has garnered heightened attention in syntactic complexity research is DD. DD offers an index for evaluating the density or dispersion of grammatical connections throughout a text. Research suggests that dependency distance minimization (DDM) reflects a universal cognitive pressure for efficient human information processing and linguistic production (<xref ref-type="bibr" rid="ref11">Futrell et al., 2015</xref>). English writers have been found to prefer syntactic structures with shorter dependencies to reduce integration difficulty and yield more locally coherent sentences (<xref ref-type="bibr" rid="ref45">Temperley, 2007</xref>). However, cross-linguistic differences have also been observed, with head-final languages like Japanese, Korean, and Turkish showing greater distances attributable to word order variation (<xref ref-type="bibr" rid="ref11">Futrell et al., 2015</xref>). Chinese and English show different dynamic valency of words and syntactic dependency structures (<xref ref-type="bibr" rid="ref35">Lu et al., 2018</xref>). <xref ref-type="bibr" rid="ref31">Liu et al. (2009)</xref> also found that Chinese shows quite different features in dependency relations, with its dependencies tending to be governor-final and mean dependency distance (MDD) being much higher than languages like English, German, and Japanese. While research has examined DDs in native language writing, fewer studies have investigated DDs in L2 academic writing.</p>
</sec>
<sec id="sec3">
<label>1.2</label>
<title>DD optimization in L1 academic writing</title>
<p>Research consistently shows a strong tendency for compact, local syntactic structures in academic writing by L1 writers, which is argued to reflect pressures for efficient linguistic processing and production (<xref ref-type="bibr" rid="ref32">Liu et al., 2017</xref>). An early study by <xref ref-type="bibr" rid="ref45">Temperley (2007)</xref> analyzed DDs in the Wall Street Journal portion of the Penn Treebank. It is found that writers favor structures with shorter dependencies, which is evidenced by their preference for short left-branching constituents. Temperley argued that writers optimize and minimize dependency lengths to yield more incrementally interpretable sentences to facilitate comprehension. <xref ref-type="bibr" rid="ref11">Futrell et al. (2015)</xref> also concluded that DDM is a universal characteristic across human languages, suggesting that variation in language can be explained by the general properties of human information processing. The authors argue that minimizing DDs enhances the efficiency of parsing and producing natural language, reducing integration costs and enabling more efficient packing of information into sentences. <xref ref-type="bibr" rid="ref36">Lu and Liu (2020)</xref> also discovered a tendency of DDM within noun phrases, potentially due to limitations in human working memory capacity.</p>
<p>However, cross-linguistic differences have also been observed, attributable to syntactic variations across languages. <xref ref-type="bibr" rid="ref11">Futrell et al. (2015)</xref> found that head-final languages like Japanese, Turkish, and Korean show much less DDM than head-initial languages like English, Italian, and Indonesian. Temperley and Gildea also found great differences across languages in DD. Their study confirms that DDM serves as an important factor in language structure and cognition, which is evidenced by the fact that writers and speakers tend to prefer structures that reduce dependency length when a language allows for different orderings of constituents (<xref ref-type="bibr" rid="ref47">Temperley and Gildea, 2018</xref>). Nonetheless, <xref ref-type="bibr" rid="ref11">Futrell et al. (2015)</xref> argue that DDM remains a universal quantitative property, as overall DDs were substantially shorter than random baselines (benchmarks created by randomly reorganizing the head word and its dependents in dependency trees, without following any specific linguistic word order rules), across all 37 diverse languages in their study. They contend that despite the structural variations among languages influencing their DDs, there is a universal aim in all languages to minimize DDs for the sake of efficiency, within the bounds of their structural limitations. The following example demonstrates the impact of syntactic variations on DD, with the specifics of the dependency relations delineated in <xref ref-type="table" rid="tab1">Table 1</xref>.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>Dependency relations of Examples 1a and b.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Dependent ID</th>
<th align="left" valign="top">Token</th>
<th align="left" valign="top">Part of Speech</th>
<th align="center" valign="top">Governor ID</th>
<th align="left" valign="top">Dependency Relation</th>
<th align="center" valign="top">Dependency Distance</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">1</td>
<td align="left" valign="middle">John</td>
<td align="left" valign="middle">NNP</td>
<td align="center" valign="top">2</td>
<td align="left" valign="middle">nsubj</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">2</td>
<td align="left" valign="middle">threw</td>
<td align="left" valign="middle">VBD</td>
<td align="center" valign="top">0</td>
<td align="left" valign="middle">ROOT</td>
<td align="center" valign="middle">/</td>
</tr>
<tr>
<td align="left" valign="middle">3</td>
<td align="left" valign="middle">out</td>
<td align="left" valign="middle">RP</td>
<td align="center" valign="top">2</td>
<td align="left" valign="middle">compound: prt</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">4</td>
<td align="left" valign="middle">the</td>
<td align="left" valign="middle">DT</td>
<td align="center" valign="top">6</td>
<td align="left" valign="middle">det</td>
<td align="center" valign="middle">2</td>
</tr>
<tr>
<td align="left" valign="middle">5</td>
<td align="left" valign="middle">old</td>
<td align="left" valign="middle">JJ</td>
<td align="center" valign="top">6</td>
<td align="left" valign="middle">amod</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">6</td>
<td align="left" valign="middle">trash</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="top">2</td>
<td align="left" valign="middle">obj</td>
<td align="center" valign="middle">4</td>
</tr>
<tr>
<td align="left" valign="middle">7</td>
<td align="left" valign="middle">sitting</td>
<td align="left" valign="middle">VBG</td>
<td align="center" valign="top">6</td>
<td align="left" valign="middle">acl</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">8</td>
<td align="left" valign="middle">in</td>
<td align="left" valign="middle">IN</td>
<td align="center" valign="top">7</td>
<td align="left" valign="middle">case</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">9</td>
<td align="left" valign="middle">the</td>
<td align="left" valign="middle">DT</td>
<td align="center" valign="top">10</td>
<td align="left" valign="middle">det</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">10</td>
<td align="left" valign="middle">kitchen</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="top">8</td>
<td align="left" valign="middle">obl</td>
<td align="center" valign="middle">2</td>
</tr>
</tbody>
</table>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="middle">Dependent ID</th>
<th align="left" valign="middle">Token</th>
<th align="left" valign="middle">Part of Speech</th>
<th align="center" valign="middle">Governor ID</th>
<th align="left" valign="middle">Dependency Relation</th>
<th align="center" valign="middle">Dependency Distance</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">1</td>
<td align="left" valign="middle">John</td>
<td align="left" valign="middle">NNP</td>
<td align="center" valign="top">2</td>
<td align="left" valign="middle">nsubj</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">2</td>
<td align="left" valign="middle">threw</td>
<td align="left" valign="middle">VBD</td>
<td align="center" valign="top">0</td>
<td align="left" valign="middle">ROOT</td>
<td align="center" valign="middle">/</td>
</tr>
<tr>
<td align="left" valign="middle">3</td>
<td align="left" valign="middle">the</td>
<td align="left" valign="middle">DT</td>
<td align="center" valign="top">5</td>
<td align="left" valign="middle">det</td>
<td align="center" valign="middle">2</td>
</tr>
<tr>
<td align="left" valign="middle">4</td>
<td align="left" valign="middle">old</td>
<td align="left" valign="middle">JJ</td>
<td align="center" valign="top">5</td>
<td align="left" valign="middle">amod</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">5</td>
<td align="left" valign="middle">trash</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="top">2</td>
<td align="left" valign="middle">obj</td>
<td align="center" valign="middle">3</td>
</tr>
<tr>
<td align="left" valign="middle">6</td>
<td align="left" valign="middle">sitting</td>
<td align="left" valign="middle">VBG</td>
<td align="center" valign="top">5</td>
<td align="left" valign="middle">acl</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">7</td>
<td align="left" valign="middle">in</td>
<td align="left" valign="middle">IN</td>
<td align="center" valign="top">6</td>
<td align="left" valign="middle">case</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">8</td>
<td align="left" valign="middle">the</td>
<td align="left" valign="middle">DT</td>
<td align="center" valign="top">9</td>
<td align="left" valign="middle">det</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">9</td>
<td align="left" valign="middle">kitchen</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="top">7</td>
<td align="left" valign="middle">obl</td>
<td align="center" valign="middle">2</td>
</tr>
<tr>
<td align="left" valign="middle">10</td>
<td align="left" valign="middle">out</td>
<td align="left" valign="middle">RP</td>
<td align="center" valign="top">2</td>
<td align="left" valign="middle">compound: prt</td>
<td align="center" valign="middle">8</td>
</tr>
</tbody>
</table>
</table-wrap>
<disp-quote>
<p>Example 1. a: John threw out the old trash sitting in the kitchen. b: John threw the old trash sitting in the kitchen out (<xref ref-type="bibr" rid="ref11">Futrell et al., 2015</xref>: 10337).</p>
</disp-quote>
<p>The two sentences above convey the same concept and have identical structures, except for the positioning of the word &#x201C;out.&#x201D; However, their total and mean DDs significantly differ, at 14 vs. 20 and 1.55 vs. 2.22, respectively. The varying DDs require distinct cognitive effort and working memory capacity. Example 1b where &#x201C;out&#x201D; is moved to the end demands more cognitive effort and working memory. For lower cognitive effort and working memory load, Example 1a, which is free from particle movement and exemplifies DDM, is preferred.</p>
<p>In sum, research shows syntax is optimized for brevity both within and across languages to aid production and comprehension, but differences exist due to language-specific conventions.</p>
</sec>
<sec id="sec4">
<label>1.3</label>
<title>DDs in L2 academic writing</title>
<p>While numerous studies have examined L1 DDs, few studies have investigated DDs in L2 academic writing. <xref ref-type="bibr" rid="ref40">Ouyang et al. (2022)</xref> investigated the writing proficiency of beginner, intermediate, and advanced learners by DD measures. They discovered that the MDD overall is significantly effective at distinguishing between each pair of consecutive proficiency levels. <xref ref-type="bibr" rid="ref16">Hao et al. (2022)</xref> verified the application of DD and its probability distribution as syntactic indicators of English as interlanguage from the perspective of language typology. They found that the MDDs of L2 learners with different backgrounds of native language gradually approach that of the target language with the improvement of their L2 proficiency. <xref ref-type="bibr" rid="ref27">Li and Yan (2021)</xref> similarly found that MDD can serve as an effective indicator to measure the syntactic complexity of Japanese EFL learners&#x2019; interlanguage. Similarly, <xref ref-type="bibr" rid="ref15">Hao et al. (2023)</xref> found in their study that dependency parameters have universal applicability in reflecting interlanguage proficiency.</p>
<p>To date, very few studies have directly compared the DD profiles of L2 writers from different L1 backgrounds composing in the L2. <xref ref-type="bibr" rid="ref12">Gao and He (2023)</xref> examined the MDD of Ph.D. dissertation abstracts written by L1 (native English) and L2 (English as a foreign language) academic writers across language backgrounds and disciplines, finding that MDD successfully distinguishes between academic texts from various linguistic backgrounds and disciplines. They argued that the authors&#x2019; efforts to make comprehension easier for readers result in the shorter MDD observed in physics and chemistry abstracts.</p>
<p>Overall, research on L2 DDs remains limited, with very few studies comparing profiles of different L1 groups in natural academic writing tasks. In particular, few studies examined whether L2 DDs and dependency directions are influenced by L2 writers&#x2019; native language. This represents a significant gap, as investigating cross-linguistic differences can elucidate the role of L1 transfer versus universality in L2 syntactic development, with key theoretical and pedagogical implications (<xref ref-type="bibr" rid="ref39">Ortega, 2003</xref>).</p>
</sec>
<sec id="sec5">
<label>1.4</label>
<title>Transfer of syntactic features in L2 acquisition</title>
<p>In examining the impact of native language transfer in L2 acquisition, scholars have extensively investigated how syntactic features influence L2 language production. <xref ref-type="bibr" rid="ref50">Whong-Barr and Schwartz (2002)</xref> reveal that in the L2 acquisition process, children&#x2019;s mastery of English dative constructions is significantly shaped by their native linguistic backgrounds, underscoring the influence of L1 syntactic frameworks and prevalent overgeneralization patterns on their learning trajectory. <xref ref-type="bibr" rid="ref5">Chan (2004)</xref> demonstrates evidence of syntactic transfer from Chinese to English among Hong Kong Chinese ESL learners, revealing that learners often think in Chinese before writing in English, leading to interlanguage structures that closely resemble or mirror the syntactic patterns of their first language, particularly in complex target structures and among learners of lower proficiency levels. Recent advancements in second language acquisition research have introduced theories such as the Interpretability Hypothesis by <xref ref-type="bibr" rid="ref48">Tsimpli and Dimitrakopoulou (2007)</xref>, which argues that learners can acquire L2 features interpretable across syntax and other cognitive systems like semantics or pragmatics, regardless of their presence in L1; the Interface Hypothesis by <xref ref-type="bibr" rid="ref44">Sorace and Filiaci (2006)</xref>, highlighting the particular difficulties learners face with language elements that integrate syntax with semantics or discourse; and the Feature Reassembly Hypothesis by <xref ref-type="bibr" rid="ref25">Lardiere (2009)</xref>, emphasizing the primary challenge of reconfiguring L1 features to conform to the target language&#x2019;s system, often leading to substantial learning challenges, especially where the languages&#x2019; feature systems notably diverge.</p>
<p>Though these new theories have been much discussed in recent years, language transfer theory remains relevant due to its powerful explanatory capabilities, able to account for many phenomena in L2 acquisition. This study aims to explore syntactic transfer through the lens of DG. We contend that the dependency patterns of an individual&#x2019;s native language can influence those in their L2 writings. For instance, consider Chinese, which is predominantly a head-final language, and English, primarily a head-initial language. The MDD of Chinese stands at 3.662, markedly higher than English&#x2019;s MDD of 2.543. We hypothesize that this significantly larger MDD in Chinese will lead to extended MDDs in L2 English writings by Chinese learners, as a consequence of language transfer. This study will verify our hypothesis.</p>
</sec>
</sec>
<sec id="sec6">
<label>2</label>
<title>Objectives and significance</title>
<p>This study delves into how L1 backgrounds influence L2 writing, particularly focusing on DDs in academic writing. It examines whether different L1 backgrounds result in distinct DD patterns in English L2 writing, potentially due to L1 transfer, or if universal linguistic principles lead to uniform patterns across L1 groups. This inquiry aims to illuminate key debates within second language acquisition (SLA) regarding the influence of native language versus universal syntax principles. It seeks to fill significant gaps in existing research and enhance writing instruction practices by clarifying the extent of cross-linguistic influence versus universal principles in L2 writing development.</p>
<p>The outcomes of this research could provide significant implications for both SLA theory and academic writing instruction. By identifying whether L2 writers&#x2019; dependency profiles are shaped more by their L1 syntax or universal syntax norms, this study inform educational strategies&#x2014;determining if writing instruction should be tailored to specific L1 backgrounds or aligned with broader, universal writing strategies. These insights will guide educators on whether to prioritize language-specific strategies or general methods to help L2 writers reach native-like proficiency.</p>
</sec>
<sec id="sec7">
<label>3</label>
<title>Theoretical framework</title>
<p>This study investigates the effect of native language on English L2 academic writing, employing quantitative analysis of dependency grammar (DG). DG, serving as a theoretical linguistic framework, delineates language structure by scrutinizing the relationships among its components. These relationships, known as dependencies, are asymmetrical connections between two constituents of a sentence, typically words, where one assumes the role of the governor or head, and the other, the dependent or modifier (<xref ref-type="bibr" rid="ref10">Fraser, 1994</xref>). In DG, DD and dependency direction, often utilized as variables in linguistic studies, serve as two critical indices for quantitative analysis. DD (<xref ref-type="bibr" rid="ref28">Liu, 2007</xref>, <xref ref-type="bibr" rid="ref29">2008</xref>; <xref ref-type="bibr" rid="ref32">Liu et al., 2017</xref>), also known as dependency length (<xref ref-type="bibr" rid="ref45">Temperley, 2007</xref>, <xref ref-type="bibr" rid="ref46">2008</xref>; <xref ref-type="bibr" rid="ref14">Gildea and Temperley, 2010</xref>; <xref ref-type="bibr" rid="ref11">Futrell et al., 2015</xref>; <xref ref-type="bibr" rid="ref47">Temperley and Gildea, 2018</xref>), refers to the linear positional difference between two words within a sentence serving as governor and dependent (<xref ref-type="bibr" rid="ref19">Hudson, 1995</xref>, <xref ref-type="bibr" rid="ref20">2010</xref>; <xref ref-type="bibr" rid="ref31">Liu et al., 2009</xref>). It is measured by the number of intervening words between dependents and their governors (<xref ref-type="bibr" rid="ref19">Hudson, 1995</xref>). For any dependency relation between two words <italic>Wx</italic> and <italic>Wy</italic>, if <italic>x</italic> is the governor and <italic>y</italic> is its dependent, their DD equals the difference <italic>x</italic>&#x2009;&#x2212;&#x2009;<italic>y</italic>; thus, adjacent words have a DD of 1. A positive distance signifies that the governor follows the dependent, whereas a negative distance indicates the governor precedes the dependent. Nevertheless, for the calculation of MDD, the absolute value of DD is used. The MDD for a sentence can be determined using the equation below:</p>
<disp-formula id="EQ1">
<label>(1)</label>
<mml:math id="M1">
<mml:mrow>
<mml:mi mathvariant="normal">MDD</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:mi mathvariant="normal">sentence</mml:mi>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mn>1</mml:mn>
<mml:mrow>
<mml:mi>n</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
</mml:mfrac>
<mml:munderover>
<mml:mstyle displaystyle="true">
<mml:mo>&#x2211;</mml:mo>
</mml:mstyle>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>=</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mi>n</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
</mml:munderover>
<mml:mo>&#x2223;</mml:mo>
<mml:mi>D</mml:mi>
<mml:msub>
<mml:mi>D</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2223;</mml:mo>
</mml:mrow>
</mml:math>
</disp-formula>
<p>where <italic>n</italic> represents the total number of words in a sentence and DD<italic>i</italic> indicates the DD of the <italic>i</italic>-th syntactic relation within the sentence (see <xref ref-type="bibr" rid="ref31">Liu et al., 2009</xref>: 166). Typically, there exists one word in each sentence that does not have a governor. This word is termed the root verb, and its DD is considered zero.</p>
<disp-formula id="EQ2">
<label>(2)</label>
<mml:math id="M2">
<mml:mrow>
<mml:mi mathvariant="normal">MDD</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:mi mathvariant="normal">sample</mml:mi>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mn>1</mml:mn>
<mml:mrow>
<mml:mi>n</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>s</mml:mi>
</mml:mrow>
</mml:mfrac>
<mml:munderover>
<mml:mstyle displaystyle="true">
<mml:mo>&#x2211;</mml:mo>
</mml:mstyle>
<mml:mrow>
<mml:mi>i</mml:mi>
<mml:mo>=</mml:mo>
<mml:mn>1</mml:mn>
</mml:mrow>
<mml:mrow>
<mml:mi>n</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi>s</mml:mi>
</mml:mrow>
</mml:munderover>
<mml:mo>&#x2223;</mml:mo>
<mml:mi>D</mml:mi>
<mml:msub>
<mml:mi>D</mml:mi>
<mml:mi>i</mml:mi>
</mml:msub>
<mml:mo>&#x2223;</mml:mo>
</mml:mrow>
</mml:math>
</disp-formula>
<p>where <italic>n</italic> represents the total number of words in a sample and <italic>s</italic> indicates the number of sentences in the sample (see <xref ref-type="bibr" rid="ref31">Liu et al., 2009</xref>: 166). DD<italic>i</italic> refers to the DD of the <italic>i</italic>-th syntactic relation within the sample.</p>
<p>The example below illustrates a dependency analysis. <xref ref-type="fig" rid="fig1">Figure 1</xref> lists the dependency structure of Example 2, while <xref ref-type="table" rid="tab2">Table 2</xref> details the dependency relations and distances associated with it.</p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Dependency structure of Example 2.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g001.tif"/>
</fig>
<table-wrap position="float" id="tab2">
<label>Table 2</label>
<caption>
<p>Dependency relations of Example 2.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Dependent id</th>
<th align="left" valign="top">Token</th>
<th align="left" valign="top">Part of speech</th>
<th align="center" valign="top">Governor id</th>
<th align="left" valign="top">Dependency relation</th>
<th align="center" valign="top">Dependency distance</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">1</td>
<td align="left" valign="middle">The</td>
<td align="left" valign="middle">DT</td>
<td align="center" valign="middle">2</td>
<td align="left" valign="middle">det</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">2</td>
<td align="left" valign="middle">table</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="middle">4</td>
<td align="left" valign="middle">nsubj</td>
<td align="center" valign="middle">2</td>
</tr>
<tr>
<td align="left" valign="middle">3</td>
<td align="left" valign="middle">below</td>
<td align="left" valign="middle">IN</td>
<td align="center" valign="middle">2</td>
<td align="left" valign="middle">case</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">4</td>
<td align="left" valign="middle">reported</td>
<td align="left" valign="middle">VBD</td>
<td align="center" valign="middle">0</td>
<td align="left" valign="middle">ROOT</td>
<td align="center" valign="middle">/</td>
</tr>
<tr>
<td align="left" valign="middle">5</td>
<td align="left" valign="middle">the</td>
<td align="left" valign="middle">DT</td>
<td align="center" valign="middle">7</td>
<td align="left" valign="middle">det</td>
<td align="center" valign="middle">2</td>
</tr>
<tr>
<td align="left" valign="middle">6</td>
<td align="left" valign="middle">dependency</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="middle">7</td>
<td align="left" valign="middle">compound</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">7</td>
<td align="left" valign="middle">relations</td>
<td align="left" valign="middle">NNS</td>
<td align="center" valign="middle">4</td>
<td align="left" valign="middle">obj</td>
<td align="center" valign="middle">3</td>
</tr>
<tr>
<td align="left" valign="middle">8</td>
<td align="left" valign="middle">and</td>
<td align="left" valign="middle">CC</td>
<td align="center" valign="middle">10</td>
<td align="left" valign="middle">cc</td>
<td align="center" valign="middle">2</td>
</tr>
<tr>
<td align="left" valign="middle">9</td>
<td align="left" valign="middle">dependency</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="middle">10</td>
<td align="left" valign="middle">compound</td>
<td align="center" valign="middle">1</td>
</tr>
<tr>
<td align="left" valign="middle">10</td>
<td align="left" valign="middle">distance</td>
<td align="left" valign="middle">NN</td>
<td align="center" valign="middle">7</td>
<td align="left" valign="middle">conj</td>
<td align="center" valign="middle">3</td>
</tr>
<tr>
<td align="left" valign="middle">11</td>
<td align="left" valign="middle">.</td>
<td align="left" valign="middle">PUNCT</td>
<td align="center" valign="middle">4</td>
<td align="left" valign="middle">punct</td>
<td align="center" valign="middle">/</td>
</tr>
</tbody>
</table>
</table-wrap>
<disp-quote>
<p>Example 2. The table below reported the dependency relations and dependency distance.</p>
</disp-quote>
<p><xref ref-type="fig" rid="fig1">Figure 1</xref> depicts the dependency relations between governors and their dependents in the example sentence. Syntactically related word pairs are connected by labeled lines with arrows pointing from the governor to the dependent. These labels, including <italic>nsubj</italic>, <italic>det</italic>, <italic>obj</italic>, <italic>conj</italic>, and <italic>punct</italic>, denote the specific dependency relations between the connected words.</p>
<p>Based on <xref ref-type="disp-formula" rid="EQ1">Eq. 1</xref>, the MDD of the example is:</p>
<disp-formula id="E1">
<mml:math id="M3">
<mml:mrow>
<mml:mi mathvariant="normal">MDD</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mrow>
<mml:mo>(</mml:mo>
<mml:mrow>
<mml:mi mathvariant="normal">Example</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mn>2</mml:mn>
</mml:mrow>
<mml:mo>)</mml:mo>
</mml:mrow>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mo>&#x2223;</mml:mo>
<mml:mn>1</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>2</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>1</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>2</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>1</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>3</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>2</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>1</mml:mn>
<mml:mo>+</mml:mo>
<mml:mn>3</mml:mn>
<mml:mo>&#x2223;</mml:mo>
</mml:mrow>
<mml:mn>9</mml:mn>
</mml:mfrac>
<mml:mo>=</mml:mo>
<mml:mn>1.7778</mml:mn>
</mml:mrow>
</mml:math>
</disp-formula>
<p>As mentioned before, DD can manifest as positive or negative contingent on whether the governor precedes or succeeds its dependent, thereby indicating the direction of dependency. When the governor precedes its dependent, DD is negative, indicating a governor-initial dependency relation, otherwise positive, denoting a governor-final dependency relation. The dependency direction within a sample can be quantified by calculating the percentages of governor-initial (or head-initial) and governor-final (or head-final) relations, using the following equations:</p>
<disp-formula id="EQ3">
<label>(3)</label>
<mml:math id="M4">
<mml:mtable columnalign="left">
<mml:mtr>
<mml:mtd>
<mml:mi mathvariant="normal">Percentage of head</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="normal">final dependency</mml:mi>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi mathvariant="normal">frequencies of the head</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="normal">final dependency</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi mathvariant="normal">total number of dependencies in the treebank</mml:mi>
</mml:mrow>
</mml:mfrac>
<mml:mo>&#x00D7;</mml:mo>
<mml:mn>100</mml:mn>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:math>
</disp-formula>
<disp-formula id="EQ4">
<label>(4)</label>
<mml:math id="M5">
<mml:mtable columnalign="left">
<mml:mtr>
<mml:mtd>
<mml:mi mathvariant="normal">Percentage of head</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="normal">initial dependency</mml:mi>
</mml:mtd>
</mml:mtr>
<mml:mtr>
<mml:mtd>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mspace width="0.25em"/>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mrow>
<mml:mi mathvariant="normal">frequencies of the head</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="normal">initial dependency</mml:mi>
</mml:mrow>
<mml:mrow>
<mml:mi mathvariant="normal">total number of dependencies in the treebank</mml:mi>
</mml:mrow>
</mml:mfrac>
<mml:mo>&#x00D7;</mml:mo>
<mml:mn>100</mml:mn>
</mml:mtd>
</mml:mtr>
</mml:mtable>
</mml:math>
</disp-formula>
<p>(see <xref ref-type="bibr" rid="ref30">Liu, 2010</xref>: 1570)</p>
<p>Applying the aforementioned <xref ref-type="disp-formula" rid="EQ3">Eqs. 3</xref> and <xref ref-type="disp-formula" rid="EQ4">4</xref>, the dependency direction of Example 2 is:</p>
<disp-formula id="E2">
<mml:math id="M6">
<mml:mrow>
<mml:mi mathvariant="normal">Percentage of head</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="normal">final dependency of Example</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mn>1</mml:mn>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mn>6</mml:mn>
<mml:mn>9</mml:mn>
</mml:mfrac>
<mml:mo>&#x00D7;</mml:mo>
<mml:mn>100</mml:mn>
<mml:mo>=</mml:mo>
<mml:mn>66.7</mml:mn>
<mml:mi>%</mml:mi>
</mml:mrow>
</mml:math>
</disp-formula>
<disp-formula id="E3">
<mml:math id="M7">
<mml:mrow>
<mml:mi mathvariant="normal">Percentage of head</mml:mi>
<mml:mo>&#x2212;</mml:mo>
<mml:mi mathvariant="normal">initial dependency of Example</mml:mi>
<mml:mspace width="thickmathspace"/>
<mml:mn>1</mml:mn>
<mml:mo>=</mml:mo>
<mml:mfrac>
<mml:mn>3</mml:mn>
<mml:mn>9</mml:mn>
</mml:mfrac>
<mml:mo>&#x00D7;</mml:mo>
<mml:mn>100</mml:mn>
<mml:mo>=</mml:mo>
<mml:mn>33.3</mml:mn>
<mml:mi>%</mml:mi>
</mml:mrow>
</mml:math>
</disp-formula>
<p>Evidently, the example sentence contains substantially more head-final dependencies compared to head-initial ones, indicating that most dependents precede their governors.</p>
<p>This study primarily focuses on the overall differences or similarities in DD between L1 and L2 English, without delving into the specific types of dependencies in each language. Our aim is to investigate the broad impact of native language on L2 writing, rather than examining the nuanced differences in dependency types and their respective DDs. These finer details of dependency types and corresponding DDs will be the subject of our future research.</p>
<p>This study&#x2019;s DG analysis benefits from recent advances in natural language processing (NLP) technology, particularly in automating part-of-speech tagging and syntactic parsing. Previously, the slow and costly manual processes hindered the development of treebanks&#x2014;key resources containing tagged and parsed sentences. However, modern NLP has overcome these challenges by enabling automatic tagging and parsing via machine learning, leveraging existing treebanks. These developments have expanded the application of NLP across the humanities, situating this research within the broader trend of integrating NLP into linguistic studies.</p>
</sec>
<sec sec-type="methods" id="sec8">
<label>4</label>
<title>Methodology</title>
<sec id="sec9">
<label>4.1</label>
<title>Research questions</title>
<p>In the present study, we intend to answer the following three questions:<list list-type="order">
<list-item>
<p>Is there a significant difference in DD between English L2 academic writings and native academic writings?</p>
</list-item>
<list-item>
<p>Is there a significant difference in dependency direction between English L2 academic writings and native academic writings?</p>
</list-item>
<list-item>
<p>Is the DD of L2 academic writings influenced by native languages?</p>
</list-item>
</list></p>
</sec>
<sec id="sec10">
<label>4.2</label>
<title>Data collection</title>
<p>The data for the present study was sourced from Scopus, selected for its status as the most extensive academic database and its established reliability as a data collection source across numerous studies (<xref ref-type="bibr" rid="ref7">Crosthwaite et al., 2022</xref>, <xref ref-type="bibr" rid="ref8">2023</xref>; <xref ref-type="bibr" rid="ref52">Zakaria and Aryadoust, 2023</xref>). Articles were chosen from the disciplines of Arts and Humanities and Social Sciences, limiting the selection to the &#x201C;article&#x201D; type while excluding &#x201C;book review&#x201D; and &#x201C;book chapter&#x201D; types, with an additional restriction for language to English only. In exporting the articles, we included the complete metadata for each article. Altogether over 2.65 million articles were extracted and stored in CSV files, with each metadata item allocated to a distinct column. Articles originating from various countries were segregated into separate CSV files, amounting to 178 files in total. The methodology for cleaning and processing the raw data is outlined in the following section.</p>
</sec>
<sec id="sec11">
<label>4.3</label>
<title>Data cleaning and processing</title>
<p>We cleaned and processed the raw data using the following procedures. First, we wrote an R script to extract the &#x201C;abstract&#x201D; and &#x201C;affiliation&#x201D; columns. Second, we extracted the country names from the affiliation column, removed rows in the abstract column where no abstract was available using an R script, and manually checked the rows in the affiliation column where the country names were not available. Following the cleaning process, we obtained more than 2.22 million abstracts, totaling over 408.9 million tokens. Next, we calculated the dependency distance and direction for each abstract in each CSV file based on <xref ref-type="disp-formula" rid="E2">Eq. 2</xref>. This calculation was performed in Python, utilizing the Stanford CoreNLP package version 4.5.5. This package was selected due to its strong performance and established reliability as an NLP tool, as evidenced by various studies (<xref ref-type="bibr" rid="ref38">Manning et al., 2014</xref>; <xref ref-type="bibr" rid="ref2">Bl&#x0161;t&#x00E1;k and Rozinajov&#x00E1;, 2022</xref>; <xref ref-type="bibr" rid="ref17">Hashemi-Namin et al., 2023</xref>; <xref ref-type="bibr" rid="ref18">He and Ang, 2023</xref>).</p>
<p>To categorize the dataset into L2 and native writings, we implemented a two-phase approach. In the first phase, we classified the abstracts according to the authors&#x2019; country of origin, identifying writings from the United Kingdom, United States, Australia, Canada, South Africa, New Zealand, Ireland, Bermuda, Jamaica, Trinidad and Tobago, Guyana, Barbados, and the Bahamas as native. In contrast, abstracts originating from any other country were designated as L2 writings. However, using affiliations to distinguish L2 from native writings is not entirely reliable, as L2 writers may study or work in countries where English is the primary language. To enhance the accuracy of classifying L2 and native writings, we introduced a second step involving a Python script that utilizes the nationalize.io API. This web-based service predicts nationalities from names using a vast database of names linked to their corresponding countries. We then compared the nationality predictions from nationalize.io with the initial phase&#x2019;s results. In cases of discrepancies between the two sets of results, which were infrequent, we performed manual verification by consulting online sources to ascertain the authors&#x2019; nationalities. Through this dual-step approach, we significantly improved the precision of our classification between L2 and native writings. Despite this thorough double-check, there might still be a small number of cases where the classification was not accurate. However, given the vast size of our dataset, totaling over 2.22 million entries, these few discrepancies are unlikely to significantly impact our overall findings. Besides, other factors might affect the quality of L2 writing. The experiences of L2 writers, such as studying abroad or having their work edited by native speakers, could contribute to the subtleties of their writing. Nevertheless, we maintain that these factors do not substantially alter the fundamental linguistic characteristics of L2 writings. While there may be exceptional instances where they do, these are not expected to cause major deviations in our overall findings. Future research could consider incorporating these variables into their study designs to further enrich and complement our findings.</p>
</sec>
</sec>
<sec sec-type="results" id="sec12">
<label>5</label>
<title>Results</title>
<sec id="sec13">
<label>5.1</label>
<title>Overall descriptive statistics</title>
<p>The table below presents the overall descriptive statistics of the data used in this study.</p>
<p>The overall descriptive statistics show that L2 writings have longer MDDs and higher percentages of governor-final dependencies (DDI) than native writings for both datasets. As the size of the datasets reached beyond the limit of Shapiro&#x2013;Wilk and Student&#x2019;s <italic>t</italic>-test, both of which require a sample size below 5,000, we used the Anderson-Darling normality test and Kolmogorov&#x2013;Smirnov test to compare the native writings and L2 writings. The results show that native writings and L2 writings are significantly different in both MDD and DDI for both datasets (<italic>p</italic>&#x2009;&#x003C;&#x2009;0.0001), indicating that L2 writings have significantly longer MDDs and higher percentages of governor-final dependencies than native writings.</p>
<p>Based on the descriptive statistics, we can offer a positive answer to our research questions 1 and 2 regarding whether there is a significant difference in dependency distance and direction between English L1 and L2 academic writings. There is a significant difference in the MDD and DDI of the two groups. Yet, it is not sure whether the significant difference is influenced by native language transfer. In the next section, we will discuss it based on more detailed results.</p>
</sec>
<sec id="sec14">
<label>5.2</label>
<title>MDD and DDI of English L2 academic writings with different language backgrounds</title>
<p>In our datasets, the sample size of some countries is very small (for example, Gambia and Guinea). Thus, we only selected those countries with a sample size of 500 and above. <xref ref-type="fig" rid="fig2">Figures 2</xref>, <xref ref-type="fig" rid="fig3">3</xref> report the dependency distances of the selected countries in the Arts and Humanities group and the Social Sciences group, respectively. (Detailed reports of the MDD and DDI are available upon request).</p>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>MDD of samples.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g002.tif"/>
</fig>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>DDI of samples.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g003.tif"/>
</fig>
<p>In <xref ref-type="table" rid="tab3">Table 3</xref>, MDD(sd) is the standard deviation of MDD, DDI is the mean ratio of dependency relations where governors are preceded by their dependents, in other words, DDI is mean governor-final ratio, and DDI(sd) is the standard deviation of DDI. The results in both sub-datasets show that the MDDs of both English native and L2 academic writings are much longer than that of the MDD of English (2.543) according to <xref ref-type="bibr" rid="ref29">Liu (2008)</xref>. Liu&#x2019;s calculation of the MDD of English is based on news texts. Our study examines the MDD of academic texts. News texts and academic texts are different genres showing different linguistic features. Their differences in MDD show that genre is a factor that affects MDD, which is partially in line with the findings of <xref ref-type="bibr" rid="ref49">Wang and Liu (2017)</xref>. According to their study, genre affects dependency distance and direction significantly, but the effect is very small. They hold that &#x201C;dependency distance is primarily determined by universal cognitive factors rather than genre-specific stylistic factors&#x201D; (<xref ref-type="bibr" rid="ref49">Wang and Liu, 2017</xref>: 135). Yet, in our study, we find that English native academic writings have an MDD of 2.9, much larger than the MDD of English news texts, which is 2.543 (<xref ref-type="bibr" rid="ref29">Liu, 2008</xref>). In <xref ref-type="bibr" rid="ref49">Wang and Liu&#x2019;s (2017)</xref> study, the ratio of dependency relations where governors precede their dependents is between 46 and 51%, while in our study, this ratio is around 33% as the ratio of governors following dependents is around 68%. This again shows that the genre of English academic writings has a much higher ratio of governor-final dependencies. Such a finding indicates that genre has a significant influence on MDD, at least in terms of the genre of academic writing. Yet, as we only examined one genre, it is not safe to claim that genre has a large effect on its influence over DD, which needs further investigation with samples from different genres.</p>
<table-wrap position="float" id="tab3">
<label>Table 3</label>
<caption>
<p>Overall descriptive statistics of the data.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top" colspan="2">Dataset</th>
<th align="center" valign="top">Number of abstracts</th>
<th align="center" valign="top">Tokens (total)</th>
<th align="center" valign="top">Tokens (mean)</th>
<th align="center" valign="top">MDD</th>
<th align="center" valign="top">MDD (sd)</th>
<th align="center" valign="top">DDI</th>
<th align="center" valign="top">DDI (sd)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="3">Social Sciences</td>
<td align="left" valign="middle">All</td>
<td align="center" valign="middle">1,360,291</td>
<td align="center" valign="middle">249,746,453</td>
<td align="center" valign="middle">183.6</td>
<td align="center" valign="middle">2.9081</td>
<td align="center" valign="middle">0.3542</td>
<td align="center" valign="middle">0.6805</td>
<td align="center" valign="middle">0.0365</td>
</tr>
<tr>
<td align="left" valign="middle">Native writings</td>
<td align="center" valign="middle">1,046,179</td>
<td align="center" valign="middle">185,956,693</td>
<td align="center" valign="middle">177.7</td>
<td align="center" valign="middle">2.8950</td>
<td align="center" valign="middle">0.3505</td>
<td align="center" valign="middle">0.6782</td>
<td align="center" valign="middle">0.0363</td>
</tr>
<tr>
<td align="left" valign="middle">L2 writings</td>
<td align="center" valign="middle">314,112</td>
<td align="center" valign="middle">63,789,760</td>
<td align="center" valign="middle">203.1</td>
<td align="center" valign="middle">2.9517</td>
<td align="center" valign="middle">0.3628</td>
<td align="center" valign="middle">0.6887</td>
<td align="center" valign="middle">0.0360</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="3">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">All</td>
<td align="center" valign="middle">860,018</td>
<td align="center" valign="middle">159,209,295</td>
<td align="center" valign="middle">185.1</td>
<td align="center" valign="middle">2.9159</td>
<td align="center" valign="middle">0.3686</td>
<td align="center" valign="middle">0.6823</td>
<td align="center" valign="middle">0.1238</td>
</tr>
<tr>
<td align="left" valign="middle">Native writings</td>
<td align="center" valign="middle">457,919</td>
<td align="center" valign="middle">81,836,408</td>
<td align="center" valign="middle">178.7</td>
<td align="center" valign="middle">2.9055</td>
<td align="center" valign="middle">0.3819</td>
<td align="center" valign="middle">0.6804</td>
<td align="center" valign="middle">0.0448</td>
</tr>
<tr>
<td align="left" valign="middle">L2 writings</td>
<td align="center" valign="middle">402,099</td>
<td align="center" valign="middle">77,372,887</td>
<td align="center" valign="middle">192.4</td>
<td align="center" valign="middle">2.9278</td>
<td align="center" valign="middle">0.3525</td>
<td align="center" valign="middle">0.6843</td>
<td align="center" valign="middle">0.1746</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>We proceeded to make Mann&#x2013;Whitney U tests between native writings and L2 writings with different language backgrounds for the two sub-datasets. Mann Whitney U test is chosen because the native group has a very large sample size and is not in a normal distribution. As the results of Mann&#x2013;Whitney U tests include many pairs of comparison, which takes up much space, they are not reported here and are available upon request. The results reveal that English native academic writings are significantly different in both MDD and DDI from English L2 academic writings of different language backgrounds for both sub-disciplines, as the <italic>p</italic> values are all below the significance level (0.05), with large effect sizes (<italic>R</italic>&#x2009;&#x003E;&#x2009;=0.8) for most pairs. This finding further confirms the result reported previously in <xref ref-type="table" rid="tab3">Table 3</xref> where English native academic writings are found to be significantly different from English L2 academic writings on the whole.</p>
<p>As the Mann&#x2013;Whitney U test examines whether two samples come from the same population, but does not reveal the correlation between variables, we did correlation analyses to find whether the differences in MDD and DDI are related to the nature of the samples, that is, native academic writings or L2 academic writings, to explore whether the language backgrounds of English L2 academic writings affect their dependency distances and dependency directions. We made a binomial logistic regression in R by the basic function <italic>glm</italic> with the two levels of the abstract type, Native vs. L2 as the response variable and MDD and DDI as predictor variables. Besides, we did another two analyses, linear regression analysis in R by the basic function <italic>lm</italic> and correlation analysis in R by the basic function <italic>cor.test</italic>. The three analyses are made for mutual corroboration.</p>
<p>The results of the binomial logistic regression in <xref ref-type="table" rid="tab4">Table 4</xref> show a significant correlation between article type and MDD and DDI for both sub-datasets, with <italic>p</italic> values below the significance level (0.05). Since the article type is a binary categorical variable, native vs. L2, the strong correlation indicates that whether the article type is native or L2 has a significant influence over MDD and DDI.</p>
<table-wrap position="float" id="tab4">
<label>Table 4</label>
<caption>
<p>Binomial logistic regression analysis results.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Discipline</th>
<th align="left" valign="top">Variable</th>
<th align="center" valign="top">Estimate</th>
<th align="center" valign="top">Std. Error</th>
<th align="center" valign="top"><italic>z</italic>.value</th>
<th align="center" valign="top">p</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="3">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">(Intercept)</td>
<td align="center" valign="middle">&#x2212;3.2263</td>
<td align="center" valign="middle">0.0439</td>
<td align="center" valign="middle">&#x2212;73.4981</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">MDD</td>
<td align="center" valign="middle">0.0429</td>
<td align="center" valign="middle">0.0062</td>
<td align="center" valign="middle">6.8887</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">DDI</td>
<td align="center" valign="middle">4.3710</td>
<td align="center" valign="middle">0.0607</td>
<td align="center" valign="middle">71.9540</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="3">Social Sciences</td>
<td align="left" valign="middle">(Intercept)</td>
<td align="center" valign="middle">&#x2212;7.4244</td>
<td align="center" valign="middle">0.0417</td>
<td align="center" valign="middle">&#x2212;177.8757</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">MDD</td>
<td align="center" valign="middle">0.3766</td>
<td align="center" valign="middle">0.0057</td>
<td align="center" valign="middle">66.2590</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">DDI</td>
<td align="center" valign="middle">7.4941</td>
<td align="center" valign="middle">0.0575</td>
<td align="center" valign="middle">130.2762</td>
<td align="center" valign="middle">0.0000</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The results of linear regression in <xref ref-type="table" rid="tab5">Table 5</xref> and correlation analyses in <xref ref-type="table" rid="tab6">Table 6</xref> both confirm the correlation between article type and MDD and DDI, with <italic>p</italic> values much lower than the significance level (0.05). As the three correlation-related analyses all show a significant influence of article type on MDD and DDI, it could be claimed that on the whole there is an effect of native language transfer on the MDD and DDI of English L2 academic writing. Currently, no study has been found to examine the MDD and DDI of all languages in the world due to various reasons, though some studies have investigated the DDs of some languages, such as <xref ref-type="bibr" rid="ref6">Chen and Gerdes (2020)</xref>, <xref ref-type="bibr" rid="ref22">Jing and Liu (2017)</xref>, <xref ref-type="bibr" rid="ref11">Futrell et al. (2015)</xref> and <xref ref-type="bibr" rid="ref29">Liu (2008)</xref>. However, examining studies such as <xref ref-type="bibr" rid="ref29">Liu (2008)</xref> reveals a trend: the greater the MDD in the background language of English L2 academic writings, the longer the MDD tends to be in English L2 academic writings themselves. For example, in the data of Social Sciences, L2 writings with Chinese as their background language have a much higher DD than the English native ones, with their MDD being 2.9896 vs. 2.8950, compared to the MDD of original Chinese and English, which is 3.662 vs. 2.543 (<xref ref-type="bibr" rid="ref29">Liu, 2008</xref>). It is the same, for instance, with Hungarian, German, and Spanish, 2.9500 vs. 2.8950 and 3.446 vs. 2.543, 2.9464 vs. 2.8950 and 3.353 vs. 2.543, 2.9400 vs. 2.8950 and 2.665 vs. 2.543.</p>
<table-wrap position="float" id="tab5">
<label>Table 5</label>
<caption>
<p>Linear regression analysis results.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Discipline</th>
<th align="left" valign="top">Variable</th>
<th align="center" valign="top">Estimate</th>
<th align="center" valign="top">Std. Error</th>
<th align="center" valign="top"><italic>t</italic>.value</th>
<th align="center" valign="top">p</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="3">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">(Intercept)</td>
<td align="center" valign="middle">&#x2212;0.2974</td>
<td align="center" valign="middle">0.0108</td>
<td align="center" valign="middle">&#x2212;27.6003</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">MDD</td>
<td align="center" valign="middle">0.0107</td>
<td align="center" valign="middle">0.0015</td>
<td align="center" valign="middle">6.9364</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">DDI</td>
<td align="center" valign="middle">1.0798</td>
<td align="center" valign="middle">0.0149</td>
<td align="center" valign="middle">72.4425</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="3">Social Sciences</td>
<td align="left" valign="middle">(Intercept)</td>
<td align="center" valign="middle">&#x2212;0.8485</td>
<td align="center" valign="middle">0.0071</td>
<td align="center" valign="middle">&#x2212;119.8052</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">MDD</td>
<td align="center" valign="middle">0.0676</td>
<td align="center" valign="middle">0.0010</td>
<td align="center" valign="middle">66.5099</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">DDI</td>
<td align="center" valign="middle">1.2973</td>
<td align="center" valign="middle">0.0099</td>
<td align="center" valign="middle">131.5301</td>
<td align="center" valign="middle">0.0000</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="tab6">
<label>Table 6</label>
<caption>
<p>Correlation analysis results.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Discipline</th>
<th align="left" valign="top">Variable</th>
<th align="center" valign="top">Cor_Coefficient</th>
<th align="center" valign="top">p</th>
<th align="center" valign="top">t</th>
<th align="center" valign="top">CI_Lower</th>
<th align="center" valign="top">CI_Upper</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="2">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">MDD</td>
<td align="center" valign="middle">0.0139</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">12.8731</td>
<td align="center" valign="middle">0.0118</td>
<td align="center" valign="middle">0.0160</td>
</tr>
<tr>
<td align="left" valign="middle">DDI</td>
<td align="center" valign="middle">0.0789</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">73.2546</td>
<td align="center" valign="middle">0.0768</td>
<td align="center" valign="middle">0.0810</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="2">Social Sciences</td>
<td align="left" valign="middle">MDD</td>
<td align="center" valign="middle">0.0674</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">78.8302</td>
<td align="center" valign="middle">0.0658</td>
<td align="center" valign="middle">0.0691</td>
</tr>
<tr>
<td align="left" valign="middle">DDI</td>
<td align="center" valign="middle">0.1177</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">138.2306</td>
<td align="center" valign="middle">0.1160</td>
<td align="center" valign="middle">0.1194</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="sec15">
<label>5.3</label>
<title>MDD and DDI of English L2 academic writings from different language families</title>
<p>To further confirm or examine the effect of native language transfer on the MDD and DDI of English L2 academic writing, we categorized the language backgrounds into several groups based on the classification of language families by <xref ref-type="bibr" rid="ref24">Katzner (2002)</xref> and <xref ref-type="bibr" rid="ref3">Brown and Ogilvie (2009)</xref>. Then, we calculated the MDD and DDI of the English L2 academic writings of different language background groups and made Mann&#x2013;Whitney U test.</p>
<p><xref ref-type="table" rid="tab7">Tables 7</xref>, <xref ref-type="table" rid="tab8">8</xref> report the MDD and DDI of the samples grouped by language family, which are accompanied by <xref ref-type="fig" rid="fig4">Figures 4</xref>, <xref ref-type="fig" rid="fig5">5</xref> for better visualization. The results show that English native academic writings are significantly different in MDD and DDI from L2 academic writings with different language family backgrounds, which can be confirmed by the results of Mann&#x2013;Whitney U test reported in <xref ref-type="table" rid="tab9">Table 9</xref>, as the <italic>p</italic> values of the comparison of most pairs between English native academic writings and L2 academic writings are below significance level (0.05). One exception is the pair of Indo-European_Native vs. Pidgin-Creole in Arts &#x0026; Humanities (<italic>p</italic>&#x2009;=&#x2009;0.0503), showing no significant difference. This insignificance arises probably because pidgins and creoles are hybrid languages formed by the blending of different languages. For example, a pidgin language is one with vocabulary &#x201C;of English, French, Spanish, or Portuguese origin&#x201D; (<xref ref-type="bibr" rid="ref24">Katzner, 2002</xref>: 32). As a result, it may share syntactical and lexical features with English, which in turn can influence the pidgin speakers&#x2019; English L2 academic writing. On the whole, a significant difference exists in MDD and DDI between English native academic writings and English L2 academic writings.</p>
<table-wrap position="float" id="tab7">
<label>Table 7</label>
<caption>
<p>MDD and DDI of samples from arts &#x0026; humanities grouped by language family.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Language Family</th>
<th align="center" valign="top">Number of abstracts</th>
<th align="center" valign="top">Tokens (total)</th>
<th align="center" valign="top">Tokens (mean)</th>
<th align="center" valign="top">MDD</th>
<th align="center" valign="top">MDD (sd)</th>
<th align="center" valign="top">DDI</th>
<th align="center" valign="top">DDI (sd)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="center" valign="middle">454,861</td>
<td align="center" valign="middle">81,825,987</td>
<td align="center" valign="middle">179.9</td>
<td align="center" valign="middle">2.9185</td>
<td align="center" valign="middle">0.3493</td>
<td align="center" valign="middle">0.6783</td>
<td align="center" valign="middle">0.0363</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European</td>
<td align="center" valign="middle">255,527</td>
<td align="center" valign="middle">48,960,196</td>
<td align="center" valign="middle">191.6</td>
<td align="center" valign="middle">2.9286</td>
<td align="center" valign="middle">0.3489</td>
<td align="center" valign="middle">0.6838</td>
<td align="center" valign="middle">0.0358</td>
</tr>
<tr>
<td align="left" valign="middle">Sino-Tibetan</td>
<td align="center" valign="middle">38,485</td>
<td align="center" valign="middle">7,262,799</td>
<td align="center" valign="middle">188.7</td>
<td align="center" valign="middle">2.9738</td>
<td align="center" valign="middle">0.3522</td>
<td align="center" valign="middle">0.6896</td>
<td align="center" valign="middle">0.0365</td>
</tr>
<tr>
<td align="left" valign="middle">Afro-Asiatic</td>
<td align="center" valign="middle">24,341</td>
<td align="center" valign="middle">4,673,168</td>
<td align="center" valign="middle">192.0</td>
<td align="center" valign="middle">2.9042</td>
<td align="center" valign="middle">0.3813</td>
<td align="center" valign="middle">0.6804</td>
<td align="center" valign="middle">0.0353</td>
</tr>
<tr>
<td align="left" valign="middle">Austronesian</td>
<td align="center" valign="middle">17,329</td>
<td align="center" valign="middle">3,611,675</td>
<td align="center" valign="middle">208.4</td>
<td align="center" valign="middle">2.8372</td>
<td align="center" valign="middle">0.3388</td>
<td align="center" valign="middle">0.6805</td>
<td align="center" valign="middle">0.0334</td>
</tr>
<tr>
<td align="left" valign="middle">Uralic</td>
<td align="center" valign="middle">13,124</td>
<td align="center" valign="middle">2,474,798</td>
<td align="center" valign="middle">188.6</td>
<td align="center" valign="middle">2.9242</td>
<td align="center" valign="middle">0.3454</td>
<td align="center" valign="middle">0.6884</td>
<td align="center" valign="middle">0.0360</td>
</tr>
<tr>
<td align="left" valign="middle">Altaic</td>
<td align="center" valign="middle">9,506</td>
<td align="center" valign="middle">1,892,020</td>
<td align="center" valign="middle">199.0</td>
<td align="center" valign="middle">2.9380</td>
<td align="center" valign="middle">0.3466</td>
<td align="center" valign="middle">0.6860</td>
<td align="center" valign="middle">0.0353</td>
</tr>
<tr>
<td align="left" valign="middle">Independent_Japanese</td>
<td align="center" valign="middle">8,379</td>
<td align="center" valign="middle">1,622,280</td>
<td align="center" valign="middle">193.6</td>
<td align="center" valign="middle">2.9616</td>
<td align="center" valign="middle">0.3553</td>
<td align="center" valign="middle">0.6885</td>
<td align="center" valign="middle">0.0366</td>
</tr>
<tr>
<td align="left" valign="middle">Niger-Congo</td>
<td align="center" valign="middle">7,124</td>
<td align="center" valign="middle">1,477,862</td>
<td align="center" valign="middle">207.4</td>
<td align="center" valign="middle">2.8946</td>
<td align="center" valign="middle">0.3479</td>
<td align="center" valign="middle">0.6751</td>
<td align="center" valign="middle">0.0331</td>
</tr>
<tr>
<td align="left" valign="middle">Independent_Korean</td>
<td align="center" valign="middle">7,112</td>
<td align="center" valign="middle">1,421,729</td>
<td align="center" valign="middle">199.9</td>
<td align="center" valign="middle">2.9538</td>
<td align="center" valign="middle">0.3402</td>
<td align="center" valign="middle">0.6887</td>
<td align="center" valign="middle">0.0344</td>
</tr>
<tr>
<td align="left" valign="middle">Tai</td>
<td align="center" valign="middle">2,757</td>
<td align="center" valign="middle">567,983</td>
<td align="center" valign="middle">206.0</td>
<td align="center" valign="middle">2.9292</td>
<td align="center" valign="middle">0.3634</td>
<td align="center" valign="middle">0.6847</td>
<td align="center" valign="middle">0.0338</td>
</tr>
<tr>
<td align="left" valign="middle">Mon-Khmer</td>
<td align="center" valign="middle">1,286</td>
<td align="center" valign="middle">249,436</td>
<td align="center" valign="middle">194.0</td>
<td align="center" valign="middle">2.9065</td>
<td align="center" valign="middle">0.3495</td>
<td align="center" valign="middle">0.6821</td>
<td align="center" valign="middle">0.0347</td>
</tr>
<tr>
<td align="left" valign="middle">Caucasian</td>
<td align="center" valign="middle">387</td>
<td align="center" valign="middle">72,489</td>
<td align="center" valign="middle">187.3</td>
<td align="center" valign="middle">2.9585</td>
<td align="center" valign="middle">0.3326</td>
<td align="center" valign="middle">0.6865</td>
<td align="center" valign="middle">0.0367</td>
</tr>
<tr>
<td align="left" valign="middle">Pidgin-Creole</td>
<td align="center" valign="middle">132</td>
<td align="center" valign="middle">26,791</td>
<td align="center" valign="middle">203.0</td>
<td align="center" valign="middle">2.9660</td>
<td align="center" valign="middle">0.3143</td>
<td align="center" valign="middle">0.6855</td>
<td align="center" valign="middle">0.0322</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="tab8">
<label>Table 8</label>
<caption>
<p>MDD and DDI of samples from social sciences grouped by language family.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Language Family</th>
<th align="center" valign="top">Number of abstracts</th>
<th align="center" valign="top">Tokens (total)</th>
<th align="center" valign="top">Tokens (mean)</th>
<th align="center" valign="top">MDD</th>
<th align="center" valign="top">MDD (sd)</th>
<th align="center" valign="top">DDI</th>
<th align="center" valign="top">DDI (sd)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="center" valign="middle">1,046,287</td>
<td align="center" valign="middle">185,979,580</td>
<td align="center" valign="middle">177.8</td>
<td align="center" valign="middle">2.8950</td>
<td align="center" valign="middle">0.3505</td>
<td align="center" valign="middle">0.6782</td>
<td align="center" valign="middle">0.0363</td>
</tr>
<tr>
<td align="left" valign="middle">Sino-Tibetan</td>
<td align="center" valign="middle">134,448</td>
<td align="center" valign="middle">26,907,979</td>
<td align="center" valign="middle">200.1</td>
<td align="center" valign="middle">2.9836</td>
<td align="center" valign="middle">0.3665</td>
<td align="center" valign="middle">0.6948</td>
<td align="center" valign="middle">0.0360</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European</td>
<td align="center" valign="middle">89,961</td>
<td align="center" valign="middle">18,757,964</td>
<td align="center" valign="middle">208.5</td>
<td align="center" valign="middle">2.9155</td>
<td align="center" valign="middle">0.3629</td>
<td align="center" valign="middle">0.6813</td>
<td align="center" valign="middle">0.0351</td>
</tr>
<tr>
<td align="left" valign="middle">Independent_Korean</td>
<td align="center" valign="middle">26,172</td>
<td align="center" valign="middle">5,210,346</td>
<td align="center" valign="middle">199.1</td>
<td align="center" valign="middle">2.9435</td>
<td align="center" valign="middle">0.3435</td>
<td align="center" valign="middle">0.6913</td>
<td align="center" valign="middle">0.0339</td>
</tr>
<tr>
<td align="left" valign="middle">Independent_Japanese</td>
<td align="center" valign="middle">23,012</td>
<td align="center" valign="middle">4,619,206</td>
<td align="center" valign="middle">200.7</td>
<td align="center" valign="middle">2.9594</td>
<td align="center" valign="middle">0.3551</td>
<td align="center" valign="middle">0.6898</td>
<td align="center" valign="middle">0.0358</td>
</tr>
<tr>
<td align="left" valign="middle">Tai</td>
<td align="center" valign="middle">9,198</td>
<td align="center" valign="middle">1,985,591</td>
<td align="center" valign="middle">215.9</td>
<td align="center" valign="middle">2.9376</td>
<td align="center" valign="middle">0.3866</td>
<td align="center" valign="middle">0.6833</td>
<td align="center" valign="middle">0.0339</td>
</tr>
<tr>
<td align="left" valign="middle">Afro-Asiatic</td>
<td align="center" valign="middle">6,266</td>
<td align="center" valign="middle">1,325,606</td>
<td align="center" valign="middle">211.6</td>
<td align="center" valign="middle">2.9232</td>
<td align="center" valign="middle">0.3597</td>
<td align="center" valign="middle">0.6812</td>
<td align="center" valign="middle">0.0350</td>
</tr>
<tr>
<td align="left" valign="middle">Mon-Khmer</td>
<td align="center" valign="middle">4,263</td>
<td align="center" valign="middle">882,930</td>
<td align="center" valign="middle">207.1</td>
<td align="center" valign="middle">2.9214</td>
<td align="center" valign="middle">0.3585</td>
<td align="center" valign="middle">0.6833</td>
<td align="center" valign="middle">0.0323</td>
</tr>
<tr>
<td align="left" valign="middle">Niger-Congo</td>
<td align="center" valign="middle">3,347</td>
<td align="center" valign="middle">744,612</td>
<td align="center" valign="middle">222.5</td>
<td align="center" valign="middle">2.9233</td>
<td align="center" valign="middle">0.3219</td>
<td align="center" valign="middle">0.6717</td>
<td align="center" valign="middle">0.0343</td>
</tr>
<tr>
<td align="left" valign="middle">Austronesian</td>
<td align="center" valign="middle">3,226</td>
<td align="center" valign="middle">686,844</td>
<td align="center" valign="middle">212.9</td>
<td align="center" valign="middle">2.9224</td>
<td align="center" valign="middle">0.3424</td>
<td align="center" valign="middle">0.6810</td>
<td align="center" valign="middle">0.0326</td>
</tr>
<tr>
<td align="left" valign="middle">Uralic</td>
<td align="center" valign="middle">2,559</td>
<td align="center" valign="middle">527,516</td>
<td align="center" valign="middle">206.1</td>
<td align="center" valign="middle">2.9255</td>
<td align="center" valign="middle">0.3339</td>
<td align="center" valign="middle">0.6816</td>
<td align="center" valign="middle">0.0369</td>
</tr>
<tr>
<td align="left" valign="middle">Altaic</td>
<td align="center" valign="middle">1,812</td>
<td align="center" valign="middle">359,683</td>
<td align="center" valign="middle">198.5</td>
<td align="center" valign="middle">2.9161</td>
<td align="center" valign="middle">0.3294</td>
<td align="center" valign="middle">0.6836</td>
<td align="center" valign="middle">0.0354</td>
</tr>
<tr>
<td align="left" valign="middle">Caucasian</td>
<td align="center" valign="middle">172</td>
<td align="center" valign="middle">35,129</td>
<td align="center" valign="middle">204.2</td>
<td align="center" valign="middle">2.9823</td>
<td align="center" valign="middle">0.3865</td>
<td align="center" valign="middle">0.6754</td>
<td align="center" valign="middle">0.0429</td>
</tr>
<tr>
<td align="left" valign="middle">Pidgin-Creole</td>
<td align="center" valign="middle">116</td>
<td align="center" valign="middle">26,465</td>
<td align="center" valign="middle">228.1</td>
<td align="center" valign="middle">2.9528</td>
<td align="center" valign="middle">0.3012</td>
<td align="center" valign="middle">0.6751</td>
<td align="center" valign="middle">0.0331</td>
</tr>
</tbody>
</table>
</table-wrap>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>MDD of samples grouped by language family.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g004.tif"/>
</fig>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>DDI of samples grouped by language family.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g005.tif"/>
</fig>
<table-wrap position="float" id="tab9">
<label>Table 9</label>
<caption>
<p>Results of Mann&#x2013;Whitney U test of MDD for samples grouped by language family.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Discipline</th>
<th align="left" valign="top">Language Family 1</th>
<th align="left" valign="top">Language Family 2</th>
<th align="center" valign="top">p</th>
<th align="center" valign="top">R</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="13">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Austronesian</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">&#x2212;0.1552</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Sino-Tibetan</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0976</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Independent_Japanese</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0807</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Indo-European</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0172</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Independent_Korean</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0653</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Niger-Congo</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">&#x2212;0.0514</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Afro-Asiatic</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">&#x2212;0.0279</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Altaic</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0333</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Caucasian</td>
<td align="center" valign="middle">0.0082</td>
<td align="center" valign="middle">0.0776</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Uralic</td>
<td align="center" valign="middle">0.0210</td>
<td align="center" valign="middle">0.0118</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Pidgin-Creole</td>
<td align="center" valign="middle">0.0503</td>
<td align="center" valign="middle">0.0984</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Mon-Khmer</td>
<td align="center" valign="middle">0.0082</td>
<td align="center" valign="middle">0.0776</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Tai</td>
<td align="center" valign="middle">0.0006</td>
<td align="center" valign="middle">&#x2212;0.0335</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="13">Social Sciences</td>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Afro-Asiatic</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0437</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Indo-European</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0333</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Niger-Congo</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0634</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Sino-Tibetan</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.1484</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Uralic</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0699</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Austronesian</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0492</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Caucasian</td>
<td align="center" valign="middle">0.0045</td>
<td align="center" valign="middle">0.1251</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Pidgin-Creole</td>
<td align="center" valign="middle">0.0172</td>
<td align="center" valign="middle">0.1277</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Independent_Japanese</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.1113</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Altaic</td>
<td align="center" valign="middle">0.0010</td>
<td align="center" valign="middle">0.0445</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Independent_Korean</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0861</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Tai</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0506</td>
</tr>
<tr>
<td align="left" valign="middle">Indo-European_Native</td>
<td align="left" valign="middle">Mon-Khmer</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0360</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Similar results in <xref ref-type="fig" rid="fig6">Figures 6</xref>, <xref ref-type="fig" rid="fig7">7</xref> are found between English native academic writings and English L2 academic writings grouped by language family group. (Detailed reports of the MDD and DDI are not presented here and are available upon request as they take much space). A significant difference arises between English native and L2 academic writings, which is confirmed by the significant p values, which is below 0.05. (Detailed reports of the Mann&#x2013;Whitney U test are available upon request to save space here).</p>
<fig position="float" id="fig6">
<label>Figure 6</label>
<caption>
<p>MDD of samples grouped by language family group.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g006.tif"/>
</fig>
<fig position="float" id="fig7">
<label>Figure 7</label>
<caption>
<p>DDI of samples grouped by language family group.</p>
</caption>
<graphic xlink:href="fpsyg-15-1384629-g007.tif"/>
</fig>
<p>We also grouped the English L2 academic writings according to whether English is regarded as an official language of the countries where these L2 writings are from. As <xref ref-type="table" rid="tab10">Table 10</xref> shows, English L2 academic writings with a background of non-English official languages have a longer MDD and higher ratio of DDI, with a significant difference (<italic>p</italic>&#x2009;&#x003C;&#x2009;0.05) from the English native academic writings as <xref ref-type="table" rid="tab11">Table 11</xref> shows. For those L2 writings with English as the official language of their countries, no significant difference (<italic>p</italic>&#x2009;=&#x2009;0.5798) from English native academic writings is found for samples from Social Sciences, though a significant difference is found for those from the Arts and Humanities.</p>
<table-wrap position="float" id="tab10">
<label>Table 10</label>
<caption>
<p>MDD and DDI of samples grouped by type of official language.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Discipline</th>
<th align="left" valign="top">Official_ Language</th>
<th align="left" valign="top">Number of abstracts</th>
<th align="left" valign="top">Tokens (total)</th>
<th align="left" valign="top">Tokens (mean)</th>
<th align="left" valign="top">MDD</th>
<th align="left" valign="top">MDD (sd)</th>
<th align="left" valign="top">DDI</th>
<th align="left" valign="top">DDI (sd)</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="3">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">454,861</td>
<td align="left" valign="middle">81,825,987</td>
<td align="left" valign="middle">179.9</td>
<td align="left" valign="middle">2.9185</td>
<td align="left" valign="middle">0.3493</td>
<td align="left" valign="middle">0.6783</td>
<td align="left" valign="middle">0.0363</td>
</tr>
<tr>
<td align="left" valign="middle">Non_English</td>
<td align="left" valign="middle">349,878</td>
<td align="left" valign="middle">67,123,651</td>
<td align="left" valign="middle">191.8</td>
<td align="left" valign="middle">2.9327</td>
<td align="left" valign="middle">0.3523</td>
<td align="left" valign="middle">0.6849</td>
<td align="left" valign="middle">0.0359</td>
</tr>
<tr>
<td align="left" valign="middle">English</td>
<td align="left" valign="middle">35,667</td>
<td align="left" valign="middle">7,201,983</td>
<td align="left" valign="middle">201.9</td>
<td align="left" valign="middle">2.8824</td>
<td align="left" valign="middle">0.3430</td>
<td align="left" valign="middle">0.6787</td>
<td align="left" valign="middle">0.0341</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="3">Social Sciences</td>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">1,046,287</td>
<td align="left" valign="middle">185,979,580</td>
<td align="left" valign="middle">177.8</td>
<td align="left" valign="middle">2.8950</td>
<td align="left" valign="middle">0.3505</td>
<td align="left" valign="middle">0.6782</td>
<td align="left" valign="middle">0.0363</td>
</tr>
<tr>
<td align="left" valign="middle">Non_English</td>
<td align="left" valign="middle">245,988</td>
<td align="left" valign="middle">49,950,349</td>
<td align="left" valign="middle">203.1</td>
<td align="left" valign="middle">2.9654</td>
<td align="left" valign="middle">0.3597</td>
<td align="left" valign="middle">0.6905</td>
<td align="left" valign="middle">0.0359</td>
</tr>
<tr>
<td align="left" valign="middle">English</td>
<td align="left" valign="middle">58,658</td>
<td align="left" valign="middle">12,141,307</td>
<td align="left" valign="middle">207.0</td>
<td align="left" valign="middle">2.8982</td>
<td align="left" valign="middle">0.3720</td>
<td align="left" valign="middle">0.6812</td>
<td align="left" valign="middle">0.0351</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap position="float" id="tab11">
<label>Table 11</label>
<caption>
<p>Results of Mann&#x2013;Whitney U test of MDD for samples grouped by official language.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Discipline</th>
<th align="left" valign="top">Language 1</th>
<th align="left" valign="top">Language 2</th>
<th align="center" valign="top">p</th>
<th align="center" valign="top">R</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="middle" rowspan="3">Arts &#x0026; Humanities</td>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">Non_English</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.0243</td>
</tr>
<tr>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">English</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">&#x2212;0.0691</td>
</tr>
<tr>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">Native</td>
<td align="center" valign="middle">1.0000</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle" rowspan="3">Social Sciences</td>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">Non_English</td>
<td align="center" valign="middle">0.0000</td>
<td align="center" valign="middle">0.1196</td>
</tr>
<tr>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">Native</td>
<td align="center" valign="middle">1.0000</td>
<td align="center" valign="middle">0.0000</td>
</tr>
<tr>
<td align="left" valign="middle">Native</td>
<td align="left" valign="middle">English</td>
<td align="center" valign="middle">0.5798</td>
<td align="center" valign="middle">&#x2212;0.0014</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec sec-type="discussion" id="sec16">
<label>6</label>
<title>Discussion</title>
<p>Native language transfer has been found in the learning and use of a second language among learners of different language backgrounds (for example, <xref ref-type="bibr" rid="ref37">Madrid, 1981</xref>; <xref ref-type="bibr" rid="ref13">Gerlach, 2017</xref>; <xref ref-type="bibr" rid="ref42">Rios Casta&#x00F1;o, 2021</xref>; <xref ref-type="bibr" rid="ref4">Chai and Bao, 2023</xref>). However, few studies have examined whether the dependency relations of English L2 academic writings are influenced by native language transfer effects, even though scholars like <xref ref-type="bibr" rid="ref43">Siu and Ho (2015)</xref> have explored the transfer of syntactic features in bilingual students. Our study of the English native and L2 academic writings within the disciplines of Arts and Humanities and Social Sciences finds a significant difference in their MDD and DDI. Significant differences in both MDD and DDI are typically observed between the native group and L2 subgroups, regardless of whether L2 academic writings are analyzed as a whole, from different language backgrounds, or from various language families. The significant difference is especially explicit when the MDDs of native English and the native languages of English L2 academic writings are significantly different. Take Chinese, Hungarian, and German, which belong to Sino-Tibetan, Uralic, and Indo-European families respectively, for example, the three languages have a much longer MDD (3.662, 3.446, 3.353) than native English (2.543). The L2 academic writings with the background of the three languages also have a much longer MDD than that of the English native academic writings (2.9896, 2.9464, and 2.9500 vs. 2.8950). For another example, when looked at from the perspective of word order typology, a significant difference is also found between native English (SVO) and the native languages (SOV) of English L2 academic writings, like Korean, which mainly falls into the type of SOV word order and is regarded by some linguists as a language in the Altaic family (<xref ref-type="bibr" rid="ref3">Brown and Ogilvie, 2009</xref>: 250). The MDD of English L2 academic writings with Korean as a native language is much longer than English native academic writings (2.9435 vs. 2.8950). English L2 academic writings with a background of native language being Spanish, which &#x201C;tends to prefer an OVS order&#x201D; (<xref ref-type="bibr" rid="ref3">Brown and Ogilvie, 2009</xref>: 884), also have a much longer MDD than that of English native academic writings (2.9400 vs. 2.8950). The MDD of Spanish is also longer than native English (2.665 vs. 2.543). Besides, in terms of DDI, English native academic writings are also significantly different from English L2 academic writings as a whole or from different language backgrounds or different language families. The majority of English L2 academic writings have a higher ratio of DDI than English native academic writings.</p>
<p>Drawing on the discussed findings, the first two research questions can now be addressed. It can be confidently stated that English L2 academic writings exhibit significant differences in MDD and DDI compared to English native academic writings. The greater the MDD in their native languages, the longer the MDD tends to be in the English L2 academic writings.</p>
<p>To address the third question, which is central to this study, we conducted regression and correlation analyses as previously discussed. These analyses investigate the potential relationship between the dependent variable (predicted) and independent (predictor) variable. In the regression analysis, the findings indicate that MDD and DDI, when used as predictor variables, successfully predict the outcomes, clearly distinguishing between native and L2 academic writings. Likewise, the results from the correlation analysis reveal that the MDD and DDI are significantly correlated to the dependent variable, distinguishing between native and L2 academic writings. Why English L2 academic writings are different in MDD and DDI from English native academic writings? Despite being published in similar or identical journals, these academic writings are authored by scholars from diverse language backgrounds: both native and non-native English speakers. For non-native English speakers, their academic writings exhibit characteristics typical of L2 texts, influenced by the phenomenon of native language transfer, as identified in previous research. The disparities in MDD and DDI between English L2 and native academic writings are likely attributed to the effect of native language transfer. This influence is particularly pronounced in L2 academic writings from background languages with a significantly longer MDD compared to native English, resulting in a substantially extended MDD in these texts.</p>
<p>However, native language transfer might not be the sole factor influencing the MDD of English L2 academic writings. For instance, we observe that English L2 academic writings with a Japanese language background exhibit a significantly longer MDD compared to English native academic writings (2.9595 vs. 2.8950), despite the fact that the MDD of Japanese itself is considerably shorter than that of native English (1.805 vs. 2.543). A plausible explanation for this phenomenon could be that, alongside native language transfer, other factors such as interlanguage interference (<xref ref-type="bibr" rid="ref1">Antoniou et al., 2011</xref>) also play a significant role. This observation suggests that native language is one of the factors influencing the dependency relations of English L2 academic writings.</p>
</sec>
<sec id="sec17">
<label>7</label>
<title>Conclusions and implications</title>
<p>Through a large dataset of English abstracts from the disciplines of Arts and Humanities and Social Sciences, the present study investigates the dependency distance and direction to examine whether native language influences the MDD of English L2 academic writings. It is found that English L2 and native academic writings differ significantly from each other in MDD and DDI. The regression and correlation analyses reveal that native language tends to be a factor influencing the MDD of English L2 academic writings. The greater the MDD of the native languages compared to that of native English, the longer the MDD in English L2 academic writings relative to English native academic writings. However, for languages with MDDs that are not significantly greater than that of native English, while the MDDs of English L2 and native academic writings differ significantly, the MDDs of English academic writings are not necessarily longer than those of English native academic writings. This observation suggests that additional factors, such as interlanguage interference, also influence the MDD of English L2 academic writings.</p>
<p>The findings could provide implications for both L2 academic writing and instruction. To increase the readability of their English academic writings, English L2 writers could try to make their writings similar to English native academic writings in MDD. For example, L2 writers from languages with significantly longer MDD than English must overcome L1 transfer effects to reduce the MDD in their English L2 academic writings. For writing instruction, the findings highlight the necessity of teaching students about the varying patterns of dependency relations between their native language and English, to make them aware of the different norms of MDD in their native language and in English. Specifically, the findings, which reveal that English L2 writings of different L1 backgrounds exhibit systematically different dependency profiles reflective of L1 transfer, underscore the value of conducting contrastive analysis between L1 and L2, as well as the importance of L1-focused instruction in academic writing pedagogy. Tailored syllabuses, targeted exercises, and native language scaffolds could be developed to help particular L1 groups reduce negative transfer effects. For example, in teaching L2 writers hailing from Chinese, Hungarian, German, and Spanish backgrounds, it is beneficial to focus on raising their consciousness to lower the MDD in English writings. Given that these languages have a significantly greater MDD than English, they have a more substantial influence on the transfer of dependency relations. This goal can be accomplished by contrasting the syntactic norms of their native languages with English, with a special emphasis on the varying patterns of dependency relations.</p>
<p>Though our study examined the MDD of English L2 writings through a large dataset, the datasets are mainly academic writings from the disciplines of Arts and Humanities and Social Sciences. Whether datasets from other disciplines or genres will yield similar or the same results is yet to be confirmed. Besides, in answering the question of whether native language influences the MDD of English L2 academic writings, we mainly rely on regression and correlation analyses. Though these analyses can reveal the causal relationship between the predictor variable MDD and DDI and the predicted variable article type (native vs. L2), it is not completely safe to conclude that native language influences the MDD of the English L2 academic writings of all language backgrounds, because there are no statistics of the MDDs of different native languages. If there are enough statistics of these MDDs in the regression and correlation analyses, it will be convincing to draw such a conclusion as more direct influence and correlation can be revealed through the analysis. Future studies could probably confirm our findings by including the MDDs of different native languages in the analysis. Besides, as our data is very large, it is unavoidable that there might be some abstracts that are not completely clean even though we have made several rounds of data cleaning. Nevertheless, the majority of our data are well-cleaned, the few unclean data do not affect the findings of our study.</p>
</sec>
<sec sec-type="data-availability" id="sec18">
<title>Data availability statement</title>
<p>The original contributions presented in the study are included in the article/supplementary material, further inquiries can be directed to the corresponding author.</p>
</sec>
<sec sec-type="author-contributions" id="sec19">
<title>Author contributions</title>
<p>YB: Conceptualization, Supervision, Writing &#x2013; review &#x0026; editing. HT: Data curation, Resources, Software, Visualization, Writing &#x2013; original draft.</p>
</sec>
</body>
<back>
<sec sec-type="funding-information" id="sec20">
<title>Funding</title>
<p>The author(s) declare that financial support was received for the research, authorship, and/or publication of this article. This research was supported by the China Postdoctoral Science Foundation (Grant No. 2023M730702); the Key Laboratory of Language Science and Multilingual Artificial Intelligence, Shanghai International Studies University, Shanghai, China (Grant No. KLSMAI-2023-OP-0008); the Center for Translation Studies of Guangdong University of Foreign Studies (Grant No. CTS202010); the Humanities and Social Sciences Funds of Department of Education of Hubei Province (Grant No. 20G012).</p>
</sec>
<sec sec-type="COI-statement" id="sec21">
<title>Conflict of interest</title>
<p>The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec sec-type="disclaimer" id="sec22">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
<ref-list>
<title>References</title>
<ref id="ref1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Antoniou</surname> <given-names>M.</given-names></name> <name><surname>Best</surname> <given-names>C. T.</given-names></name> <name><surname>Tyler</surname> <given-names>M. D.</given-names></name> <name><surname>Kroos</surname> <given-names>C.</given-names></name></person-group> (<year>2011</year>). <article-title>Inter-language interference in VOT production by l2-dominant bilinguals: asymmetries in phonetic code-switching</article-title>. <source>J. Phon.</source> <volume>39</volume>, <fpage>558</fpage>&#x2013;<lpage>570</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.wocn.2011.03.001</pub-id>, PMID: <pub-id pub-id-type="pmid">22787285</pub-id></citation></ref>
<ref id="ref2"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bl&#x0161;t&#x00E1;k</surname> <given-names>M.</given-names></name> <name><surname>Rozinajov&#x00E1;</surname> <given-names>V.</given-names></name></person-group> (<year>2022</year>). <article-title>Automatic question generation based on sentence structure analysis using machine learning approach</article-title>. <source>Nat. Lang. Eng.</source> <volume>28</volume>, <fpage>487</fpage>&#x2013;<lpage>517</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S1351324921000139</pub-id></citation></ref>
<ref id="ref3"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Brown</surname> <given-names>K.</given-names></name> <name><surname>Ogilvie</surname> <given-names>S.</given-names></name></person-group> (<year>2009</year>). <source>Concise encyclopedia of languages of the world</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Elsevier</publisher-name>.</citation></ref>
<ref id="ref4"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chai</surname> <given-names>X. S.</given-names></name> <name><surname>Bao</surname> <given-names>J.</given-names></name></person-group> (<year>2023</year>). <article-title>Linguistic distances between native languages and Chinese influence acquisition of Chinese character, vocabulary, and grammar</article-title>. <source>Front. Psychol.</source> doi: <pub-id pub-id-type="doi">10.3389/fpsyg.2022.1083574</pub-id></citation></ref>
<ref id="ref5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chan</surname> <given-names>A. Y.</given-names></name></person-group> (<year>2004</year>). <article-title>Syntactic transfer: evidence from the interlanguage of Hong Kong Chinese ESL learners</article-title>. <source>Mod. Lang. J.</source> <volume>88</volume>, <fpage>56</fpage>&#x2013;<lpage>74</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.0026-7902.2004.00218.x</pub-id></citation></ref>
<ref id="ref6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Chen</surname> <given-names>X. Y.</given-names></name> <name><surname>Gerdes</surname> <given-names>K.</given-names></name></person-group> (<year>2020</year>). <article-title>Dependency distances and their frequencies in indo-european language</article-title>. <source>J. Quant. Ling.</source> <volume>29</volume>, <fpage>106</fpage>&#x2013;<lpage>125</lpage>. doi: <pub-id pub-id-type="doi">10.1080/09296174.2020.1771135</pub-id></citation></ref>
<ref id="ref7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crosthwaite</surname> <given-names>P.</given-names></name> <name><surname>Ningrum</surname> <given-names>S.</given-names></name> <name><surname>Lee</surname> <given-names>I.</given-names></name></person-group> (<year>2022</year>). <article-title>Research trends in L2 written corrective feedback: a bibliometric analysis of three decades of Scopus-indexed research on L2 WCF</article-title>. <source>J. Second. Lang. Writ.</source> <volume>58</volume>:<fpage>100934</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.jslw.2022.100934</pub-id></citation></ref>
<ref id="ref8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Crosthwaite</surname> <given-names>P.</given-names></name> <name><surname>Ningrum</surname> <given-names>S.</given-names></name> <name><surname>Schweinberger</surname> <given-names>M.</given-names></name></person-group> (<year>2023</year>). <article-title>Research trends in corpus linguistics</article-title>. <source>Int. J. Corpus Ling.</source> <volume>28</volume>, <fpage>344</fpage>&#x2013;<lpage>377</lpage>. doi: <pub-id pub-id-type="doi">10.1075/ijcl.21072.cro</pub-id></citation></ref>
<ref id="ref9"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Fan</surname> <given-names>L.</given-names></name> <name><surname>Jiang</surname> <given-names>Y.</given-names></name></person-group> (<year>2019</year>). <article-title>Can dependency distance and direction be used to differentiate translational language from native language?</article-title> <source>Lingua</source> <volume>224</volume>, <fpage>51</fpage>&#x2013;<lpage>59</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2019.03.004</pub-id></citation></ref>
<ref id="ref10"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Fraser</surname> <given-names>N</given-names></name></person-group>. (<year>1994</year>). <article-title>Dependency grammar</article-title>, <source>Encyclopedia of language and linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Asher</surname> <given-names>R. E.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Pergamon</publisher-name>), <fpage>860</fpage>&#x2013;<lpage>864</lpage>.</citation></ref>
<ref id="ref11"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Futrell</surname> <given-names>R.</given-names></name> <name><surname>Mahowald</surname> <given-names>K.</given-names></name> <name><surname>Gibson</surname> <given-names>E.</given-names></name></person-group> (<year>2015</year>). <article-title>Large-scale evidence of dependency length minimization in 37 languages</article-title>. <source>Proc. Natl. Acad. Sci.</source> <volume>112</volume>, <fpage>10336</fpage>&#x2013;<lpage>10341</lpage>. doi: <pub-id pub-id-type="doi">10.1073/pnas.1502134112</pub-id>, PMID: <pub-id pub-id-type="pmid">26240370</pub-id></citation></ref>
<ref id="ref12"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gao</surname> <given-names>N.</given-names></name> <name><surname>He</surname> <given-names>Q.</given-names></name></person-group> (<year>2023</year>). <article-title>A corpus-based study of the dependency distance differences in English academic writing</article-title>. <source>SAGE Open</source> <volume>13</volume>:<fpage>198408</fpage>. doi: <pub-id pub-id-type="doi">10.1177/21582440231198408</pub-id></citation></ref>
<ref id="ref13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gerlach</surname> <given-names>D.</given-names></name></person-group> (<year>2017</year>). <article-title>Reading and spelling difficulties in the ELT classroom</article-title>. <source>ELT J.</source> <volume>71</volume>, <fpage>295</fpage>&#x2013;<lpage>304</lpage>. doi: <pub-id pub-id-type="doi">10.1093/elt/ccw088</pub-id></citation></ref>
<ref id="ref14"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Gildea</surname> <given-names>D.</given-names></name> <name><surname>Temperley</surname> <given-names>D.</given-names></name></person-group> (<year>2010</year>). <article-title>Do grammars minimize dependency length?</article-title> <source>Cogn. Sci.</source> <volume>34</volume>, <fpage>286</fpage>&#x2013;<lpage>310</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1551-6709.2009.01073.x</pub-id>, PMID: <pub-id pub-id-type="pmid">21564213</pub-id></citation></ref>
<ref id="ref15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hao</surname> <given-names>Y. X.</given-names></name> <name><surname>Wang</surname> <given-names>X. L.</given-names></name> <name><surname>Bin</surname> <given-names>S.</given-names></name> <name><surname>Liu</surname> <given-names>H. T.</given-names></name></person-group> (<year>2023</year>). <article-title>A probability distribution of dependencies in interlanguage</article-title>. <source>Poznan Stud. Contemp. Ling.</source> <volume>59</volume>, <fpage>65</fpage>&#x2013;<lpage>93</lpage>. doi: <pub-id pub-id-type="doi">10.1515/psicl-2022-2007</pub-id></citation></ref>
<ref id="ref16"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hao</surname> <given-names>Y. X.</given-names></name> <name><surname>Wang</surname> <given-names>X. L.</given-names></name> <name><surname>Lin</surname> <given-names>Y. N.</given-names></name></person-group> (<year>2022</year>). <article-title>Dependency distance and its probability distribution: are they the universals for measuring second language learners&#x2019; language proficiency?</article-title> <source>J. Quant. Ling.</source> <volume>29</volume>, <fpage>485</fpage>&#x2013;<lpage>509</lpage>. doi: <pub-id pub-id-type="doi">10.1080/09296174.2021.1991684</pub-id></citation></ref>
<ref id="ref17"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Hashemi-Namin</surname> <given-names>M.</given-names></name> <name><surname>Jahed-Motlagh</surname> <given-names>M.</given-names></name> <name><surname>Torkaman Rahmani</surname> <given-names>A.</given-names></name></person-group> (<year>2023</year>). <article-title>Recognition of visual scene elements from a story text in Persian natural language</article-title>. <source>Nat. Lang. Eng.</source> <volume>29</volume>, <fpage>693</fpage>&#x2013;<lpage>719</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S1351324922000390</pub-id></citation></ref>
<ref id="ref18"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>He</surname> <given-names>M. Y.</given-names></name> <name><surname>Ang</surname> <given-names>L. H.</given-names></name></person-group> (<year>2023</year>). <article-title>Profiling a microeconomics noun collocation list: a corpus-based approach</article-title>. <source>Southern Afr. Ling. Appl. Lang. Stud.</source> <volume>41</volume>, <fpage>191</fpage>&#x2013;<lpage>209</lpage>. doi: <pub-id pub-id-type="doi">10.2989/16073614.2022.2117708</pub-id></citation></ref>
<ref id="ref19"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Hudson</surname> <given-names>R. A.</given-names></name></person-group> (<year>1995</year>). Measuring syntactic difficulty. Unpublished paper. Available at: <ext-link xlink:href="https://dickhudson.com/wp-content/uploads/2013/07/Difficulty.pdf" ext-link-type="uri">https://dickhudson.com/wp-content/uploads/2013/07/Difficulty.pdf</ext-link> (2021-10-31)</citation></ref>
<ref id="ref20"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Hudson</surname> <given-names>R.</given-names></name></person-group> (<year>2010</year>). <source>An introduction to word grammar</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>.</citation></ref>
<ref id="ref21"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Jiang</surname> <given-names>J.</given-names></name> <name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2015</year>). <article-title>The effects of sentence length on dependency distance, dependency direction and the implications&#x2014;based on a parallel English&#x2013;Chinese dependency treebank</article-title>. <source>Lang. Sci.</source> <volume>50</volume>, <fpage>93</fpage>&#x2013;<lpage>104</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.langsci.2015.04.002</pub-id></citation></ref>
<ref id="ref22"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Jing</surname> <given-names>Y.</given-names></name> <name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2017</year>). <article-title>Dependency distance motifs in 21 indo-European languages</article-title>, <source>Motifs in language and text</source>. eds. <person-group person-group-type="editor"><name><surname>Liu</surname> <given-names>H.T.</given-names></name> <name><surname>Liang</surname> <given-names>J.Y.</given-names></name></person-group> (<publisher-loc>Berlin/Boston</publisher-loc>: <publisher-name>Mouton De Gruyter</publisher-name>), <fpage>133</fpage>&#x2013;<lpage>150</lpage>.</citation></ref>
<ref id="ref23"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Kaplan</surname> <given-names>R. B.</given-names></name></person-group> (<year>1966</year>). <article-title>Cultural thought patterns in intercultural education</article-title>. <source>Lang. Learn.</source> <volume>16</volume>, <fpage>1</fpage>&#x2013;<lpage>20</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1467-1770.1966.tb00804.x</pub-id></citation></ref>
<ref id="ref24"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Katzner</surname> <given-names>K.</given-names></name></person-group> (<year>2002</year>). <source>The languages of the world</source>. <publisher-loc>London and New York</publisher-loc>: <publisher-name>Routledge</publisher-name>.</citation></ref>
<ref id="ref25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lardiere</surname> <given-names>D.</given-names></name></person-group> (<year>2009</year>). <article-title>Some thoughts on the contrastive analysis of features in second language acquisition</article-title>. <source>Second. Lang. Res.</source> <volume>25</volume>, <fpage>173</fpage>&#x2013;<lpage>227</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0267658308100283</pub-id></citation></ref>
<ref id="ref26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lei</surname> <given-names>L.</given-names></name> <name><surname>Wen</surname> <given-names>J.</given-names></name></person-group> (<year>2019</year>). <article-title>Is dependency distance experiencing a process of minimization?, A diachronic study based on the state of the union addresses</article-title>. <source>Lingua</source> <volume>16</volume>:<fpage>102762</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2019.102762</pub-id></citation></ref>
<ref id="ref27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Li</surname> <given-names>W. P.</given-names></name> <name><surname>Yan</surname> <given-names>J. W.</given-names></name></person-group> (<year>2021</year>). <article-title>Probability distribution of dependency distance based on a treebank of Japanese EFL learners&#x2019; interlanguage</article-title>. <source>J. Quant. Ling.</source> <volume>28</volume>, <fpage>172</fpage>&#x2013;<lpage>186</lpage>. doi: <pub-id pub-id-type="doi">10.1080/09296174.2020.1754611</pub-id></citation></ref>
<ref id="ref28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2007</year>). <article-title>Probability distribution of dependency distance</article-title>. <source>Glottometrics.</source> <volume>15</volume>, <fpage>1</fpage>&#x2013;<lpage>12</lpage>,</citation></ref>
<ref id="ref29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2008</year>). <article-title>Dependency distance as a metric of language comprehension difficulty</article-title>. <source>J. Cogn. Sci.</source> <volume>9</volume>, <fpage>159</fpage>&#x2013;<lpage>191</lpage>. doi: <pub-id pub-id-type="doi">10.17791/jcs.2008.9.2.159</pub-id></citation></ref>
<ref id="ref30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2010</year>). <article-title>Dependency direction as a means of word-order typology: a method based on dependency treebanks</article-title>. <source>Lingua</source> <volume>120</volume>, <fpage>1567</fpage>&#x2013;<lpage>1578</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2009.10.001</pub-id></citation></ref>
<ref id="ref31"><citation citation-type="journal"><person-group person-group-type="editor"><name><surname>Liu</surname> <given-names>H.</given-names></name> <name><surname>Hudson</surname> <given-names>R.</given-names></name> <name><surname>Feng</surname> <given-names>Z.</given-names></name></person-group> (<year>2009</year>). <article-title>Using a Chinese treebank to measure dependency distance</article-title>. <source>Corpus Linguist. Linguist. Theory</source> <volume>5</volume>, <fpage>161</fpage>&#x2013;<lpage>174</lpage>. doi: <pub-id pub-id-type="doi">10.1515/CLLT.2009.007</pub-id></citation></ref>
<ref id="ref32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>H.</given-names></name> <name><surname>Xu</surname> <given-names>C.</given-names></name> <name><surname>Liang</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). <article-title>Dependency distance: a new perspective on syntactic patterns in natural languages</article-title>. <source>Phys Life Rev</source> <volume>21</volume>, <fpage>171</fpage>&#x2013;<lpage>193</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.plrev.2017.03.002</pub-id>, PMID: <pub-id pub-id-type="pmid">28624589</pub-id></citation></ref>
<ref id="ref33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Liu</surname> <given-names>X.</given-names></name> <name><surname>Zhu</surname> <given-names>H.</given-names></name> <name><surname>Lei</surname> <given-names>L.</given-names></name></person-group> (<year>2022</year>). <article-title>Dependency distance minimization: a diachronic exploration of the effects of sentence length and dependency types</article-title>. <source>Human. Soc. Sci. Commun.</source> <volume>9</volume>:<fpage>420</fpage>. doi: <pub-id pub-id-type="doi">10.1057/s41599-022-01447-3</pub-id></citation></ref>
<ref id="ref34"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lu</surname> <given-names>X.</given-names></name></person-group> (<year>2011</year>). <article-title>A corpus-based evaluation of syntactic complexity measures as indices of college-level ESL writers&#x2019; language development</article-title>. <source>TESOL Q.</source> <volume>45</volume>, <fpage>36</fpage>&#x2013;<lpage>62</lpage>. doi: <pub-id pub-id-type="doi">10.5054/tq.2011.240859</pub-id></citation></ref>
<ref id="ref35"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Lu</surname> <given-names>Q.</given-names></name> <name><surname>Lin</surname> <given-names>Y.</given-names></name> <name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2018</year>). <article-title>Dynamic valency and dependency distance</article-title>, <source>Quantitative analysis of dependency structures</source>. eds. <person-group person-group-type="editor"><name><surname>Jiang</surname> <given-names>J.Y.</given-names></name> <name><surname>Liu</surname> <given-names>H.T.</given-names></name></person-group> (<publisher-loc>Berlin</publisher-loc>: <publisher-name>De Gruyter Mouto</publisher-name>), <fpage>145</fpage>&#x2013;<lpage>166</lpage>.</citation></ref>
<ref id="ref36"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Lu</surname> <given-names>J. Y.</given-names></name> <name><surname>Liu</surname> <given-names>H. T.</given-names></name></person-group> (<year>2020</year>). <article-title>Do English noun phrases tend to minimize dependency distance?</article-title> <source>Austr. J. Ling.</source> <volume>40</volume>, <fpage>246</fpage>&#x2013;<lpage>262</lpage>. doi: <pub-id pub-id-type="doi">10.1080/07268602.2020.1789552</pub-id></citation></ref>
<ref id="ref37"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Madrid</surname> <given-names>L. D.</given-names></name></person-group> (<year>1981</year>). <source>The effect of native language transfer on the production of second language negative syntactic forms and adjective-noun sequences in relation to bilingual proficiency</source>. <publisher-loc>Santa Barbara</publisher-loc>: <publisher-name>University of California</publisher-name>.</citation></ref>
<ref id="ref38"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Manning</surname> <given-names>C.D.</given-names></name> <name><surname>Surdeanu</surname> <given-names>M.</given-names></name> <name><surname>Bauer</surname> <given-names>J.</given-names></name> <name><surname>Finkel</surname> <given-names>J.</given-names></name> <name><surname>Bethard</surname> <given-names>S.</given-names></name> <name><surname>McClosky</surname> <given-names>D.</given-names></name></person-group> (<year>2014</year>). <article-title>The Stanford CoreNLP natural language processing toolkit</article-title>, <conf-name>Proceedings of the 52nd annual meeting of the Association for Computational Linguistics: System demonstrations</conf-name>. <fpage>55</fpage>&#x2013;<lpage>60</lpage>.</citation></ref>
<ref id="ref39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ortega</surname> <given-names>L.</given-names></name></person-group> (<year>2003</year>). <article-title>Syntactic complexity measures and their relationship to l2 proficiency: a research synthesis of college-level l2 writing</article-title>. <source>Appl. Linguis.</source> <volume>24</volume>, <fpage>492</fpage>&#x2013;<lpage>518</lpage>. doi: <pub-id pub-id-type="doi">10.1093/applin/24.4.492</pub-id></citation></ref>
<ref id="ref40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ouyang</surname> <given-names>J.</given-names></name> <name><surname>Jiang</surname> <given-names>J.</given-names></name> <name><surname>Liu</surname> <given-names>H.</given-names></name></person-group> (<year>2022</year>). <article-title>Dependency distance measures in assessing l2 writing proficiency</article-title>. <source>Assess. Writ.</source> <volume>51</volume>:<fpage>100603</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.asw.2021.100603</pub-id></citation></ref>
<ref id="ref41"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Oya</surname> <given-names>M.</given-names></name></person-group> (<year>2011</year>). <article-title>Syntactic dependency distance as sentence complexity measure</article-title>, <conf-name>Proceedings of the 16th conference of Pan-Pacific Association of Applied Linguistics</conf-name>. <fpage>313</fpage>&#x2013;<lpage>316</lpage>.</citation></ref>
<ref id="ref42"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Rios Casta&#x00F1;o</surname> <given-names>A.</given-names></name></person-group> (<year>2021</year>). <source>Mother tongue interference in the English acquisition as a second language</source>: <publisher-name>Greensboro College</publisher-name>.</citation></ref>
<ref id="ref43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Siu</surname> <given-names>C. T. S.</given-names></name> <name><surname>Ho</surname> <given-names>C. S. H.</given-names></name></person-group> (<year>2015</year>). <article-title>Cross-language transfer of syntactic skills and reading comprehension among young Cantonese-English bilingual students</article-title>. <source>Read. Res. Q.</source> <volume>50</volume>, <fpage>313</fpage>&#x2013;<lpage>336</lpage>. doi: <pub-id pub-id-type="doi">10.1002/rrq.101</pub-id></citation></ref>
<ref id="ref44"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Sorace</surname> <given-names>A.</given-names></name> <name><surname>Filiaci</surname> <given-names>F.</given-names></name></person-group> (<year>2006</year>). <source>Anaphora resolution in near-native speakers of Italian. Second, Language Research</source>, vol. <volume>22</volume>, <fpage>339</fpage>&#x2013;<lpage>368</lpage>.</citation></ref>
<ref id="ref45"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Temperley</surname> <given-names>D.</given-names></name></person-group> (<year>2007</year>). <article-title>Minimization of dependency length in written English</article-title>. <source>Cognition</source> <volume>105</volume>, <fpage>300</fpage>&#x2013;<lpage>333</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.cognition.2006.09.011</pub-id>, PMID: <pub-id pub-id-type="pmid">17074312</pub-id></citation></ref>
<ref id="ref46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Temperley</surname> <given-names>D.</given-names></name></person-group> (<year>2008</year>). <article-title>Dependency-length minimization in natural and artificial languages</article-title>. <source>J. Quant. Ling.</source> <volume>15</volume>, <fpage>256</fpage>&#x2013;<lpage>282</lpage>. doi: <pub-id pub-id-type="doi">10.1080/09296170802159512</pub-id></citation></ref>
<ref id="ref47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Temperley</surname> <given-names>D.</given-names></name> <name><surname>Gildea</surname> <given-names>D.</given-names></name></person-group> (<year>2018</year>). <article-title>Minimizing syntactic dependency lengths: typological/cognitive universal?</article-title> <source>Ann. Rev. Ling.</source> <volume>4</volume>, <fpage>67</fpage>&#x2013;<lpage>80</lpage>. doi: <pub-id pub-id-type="doi">10.1146/annurev-linguistics-011817-045617</pub-id></citation></ref>
<ref id="ref48"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tsimpli</surname> <given-names>I. M.</given-names></name> <name><surname>Dimitrakopoulou</surname> <given-names>M.</given-names></name></person-group> (<year>2007</year>). <article-title>The interpretability hypothesis: evidence from WH-interrogatives in second language acquisition</article-title>. <source>Second. Lang. Res.</source> <volume>23</volume>, <fpage>215</fpage>&#x2013;<lpage>242</lpage>. doi: <pub-id pub-id-type="doi">10.1177/0267658307076546</pub-id></citation></ref>
<ref id="ref9001"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tesni&#x00E8;re</surname> <given-names>L.</given-names></name></person-group> (<year>1959</year>). <source>El&#x00E9;ments de la syntaxe structurale. Paris: Klincksieck</source>.</citation></ref>
<ref id="ref49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wang</surname> <given-names>Y. Q.</given-names></name> <name><surname>Liu</surname> <given-names>H. T.</given-names></name></person-group> (<year>2017</year>). <article-title>The effects of genre on dependency distance and dependency direction</article-title>. <source>Lang. Sci.</source> <volume>59</volume>, <fpage>135</fpage>&#x2013;<lpage>147</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.langsci.2016.09.006</pub-id></citation></ref>
<ref id="ref50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Whong-Barr</surname> <given-names>M.</given-names></name> <name><surname>Schwartz</surname> <given-names>B. D.</given-names></name></person-group> (<year>2002</year>). <article-title>Morphological and syntactic transfer in child L2 acquisition of the English dative alternation</article-title>. <source>Stud. Second. Lang. Acquis.</source> <volume>24</volume>, <fpage>579</fpage>&#x2013;<lpage>616</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0272263102004035</pub-id></citation></ref>
<ref id="ref51"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Xu</surname> <given-names>H.</given-names></name> <name><surname>Liu</surname> <given-names>K. L.</given-names></name></person-group> (<year>2023</year>). <article-title>Syntactic simplification in interpreted English: dependency distance and direction measures</article-title>. <source>Lingua</source> <volume>294</volume>:<fpage>103607</fpage>. doi: <pub-id pub-id-type="doi">10.1016/j.lingua.2023.103607</pub-id></citation></ref>
<ref id="ref52"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zakaria</surname> <given-names>A.</given-names></name> <name><surname>Aryadoust</surname> <given-names>V.</given-names></name></person-group> (<year>2023</year>). <article-title>A scientometric analysis of applied linguistics research (1970&#x2013;2022): methodology and future directions</article-title>. <source>Appl. Ling. Rev.</source> <volume>15</volume>:<fpage>e210</fpage>. doi: <pub-id pub-id-type="doi">10.1515/applirev-2022-0210</pub-id></citation></ref>
</ref-list>
</back>
</article>