<?xml version="1.0" encoding="utf-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v2.3 20070202//EN" "journalpublishing.dtd">
<article xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance" article-type="research-article" dtd-version="2.3" xml:lang="EN">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">Front. Psychol.</journal-id>
<journal-title>Frontiers in Psychology</journal-title>
<abbrev-journal-title abbrev-type="pubmed">Front. Psychol.</abbrev-journal-title>
<issn pub-type="epub">1664-1078</issn>
<publisher>
<publisher-name>Frontiers Media S.A.</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.3389/fpsyg.2023.1128461</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Psychology</subject>
<subj-group>
<subject>Original Research</subject>
</subj-group>
</subj-group>
</article-categories>
<title-group>
<article-title>Tone and word length across languages</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes"><name><surname>Wichmann</surname> <given-names>S&#x00F8;ren</given-names></name><xref rid="c001" ref-type="corresp"><sup>&#x002A;</sup></xref>
<uri xlink:href="https://loop.frontiersin.org/people/485445/overview"/>
</contrib>
</contrib-group>
<aff><institution>Cluster of Excellence ROOTS, Kiel University</institution>, <addr-line>Kiel</addr-line>, <country>Germany</country></aff>
<author-notes>
<fn id="fn0001" fn-type="edited-by">
<p>Edited by: Steven Moran, University of Neuch&#x00E2;tel, Switzerland</p>
</fn>
<fn id="fn0002" fn-type="edited-by">
<p>Reviewed by: Andrew Wedel, University of Arizona, United States; Christian Bentz, University of T&#x00FC;bingen, Germany</p>
</fn>
<corresp id="c001">&#x002A;Correspondence: S&#x00F8;ren Wichmann, <email>wichmannsoeren@gmail.com</email></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>22</day>
<month>06</month>
<year>2023</year>
</pub-date>
<pub-date pub-type="collection">
<year>2023</year>
</pub-date>
<volume>14</volume>
<elocation-id>1128461</elocation-id>
<history>
<date date-type="received">
<day>20</day>
<month>12</month>
<year>2022</year>
</date>
<date date-type="accepted">
<day>25</day>
<month>04</month>
<year>2023</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright &#x00A9; 2023 Wichmann.</copyright-statement>
<copyright-year>2023</copyright-year>
<copyright-holder>Wichmann</copyright-holder>
<license xlink:href="http://creativecommons.org/licenses/by/4.0/">
<p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.</p>
</license>
</permissions>
<abstract>
<p>The aim of this paper is to show evidence of a statistical dependency of the presence of tones on word length. Other work has made it clear that there is a strong inverse correlation between population size and word length. Here it is additionally shown that word length is coupled with tonal distinctions, languages being more likely to have such distinctions when they exhibit shorter words. It is hypothesized that the chain of causation is such that population size influences word length, which, in turn, influences the presence and number of tonal distinctions.</p>
</abstract>
<kwd-group>
<kwd>tones</kwd>
<kwd>tonogenesis</kwd>
<kwd>word length</kwd>
<kwd>linguistic diversity</kwd>
<kwd>linguistic typology</kwd>
<kwd>language size</kwd>
</kwd-group>
<counts>
<fig-count count="5"/>
<table-count count="4"/>
<equation-count count="0"/>
<ref-count count="55"/>
<page-count count="12"/>
<word-count count="9252"/>
</counts>
<custom-meta-wrap>
<custom-meta>
<meta-name>section-at-acceptance</meta-name>
<meta-value>Psychology of Language</meta-value>
</custom-meta>
</custom-meta-wrap>
</article-meta>
</front>
<body>
<sec id="sec1" sec-type="intro">
<title>Introduction</title>
<p>Previous work has investigated factors that influence word length both across meanings on a subset of the Swadesh list and across languages (<xref ref-type="bibr" rid="ref46">Wichmann and Holman, 2023</xref>). Across languages, a factor found to influence word length was population size. Aggregation across language families and six macroareas compared with similarly aggregated logs of population sizes showed an extremely strong (<italic>r</italic>&#x2009;=&#x2009;&#x2212;0.92, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.01) correlation. In other words, word length averaged across families and then across macroareas decreases as similarly averaged populations increase. This finding supports a suggestion in <xref ref-type="bibr" rid="ref50">Wichmann et al. (2011</xref>, p. 193&#x2013;194) of an existence of an inverse relationship between word length and population sizes, a suggestion which, in turn, followed an original proposal by <xref ref-type="bibr" rid="ref25">Nettle (1995</xref>, <xref ref-type="bibr" rid="ref26">1998)</xref>. The main insights from the study of <xref ref-type="bibr" rid="ref46">Wichmann and Holman (2023)</xref> may be replicated from the basic data, which have been made available online at <ext-link xlink:href="https://zenodo.org/record/6344024" ext-link-type="uri">https://zenodo.org/record/6344024</ext-link>. Data on word length was based on averages across the 40 item word lists in ASJP (<xref ref-type="bibr" rid="ref48">Wichmann et al., 2020</xref>), data which will also be used in the present study.</p>
<p>The present paper goes on to look at how presence/absence of tones as well as the number of tonal contrasts relate to mean word length. Languages with shorter words might be more susceptible to having tonal contrasts, and, beyond mere presence vs. absence, it seems worthwhile to test whether the number of tonal contrasts correlates with word length. For instance, SE Asia is famous for having a high concentration of tonal languages as well as for a tendency for languages to have monosyllabic words. In contrast, Australian languages tend to have long words and no tones. The aim of the work described in this paper is to test whether a relationship between word length and tones generalizes beyond such anecdotal cases. Research on ways that tonal distinction may emerge (tonogenesis), moreover, suggests a plausible causal connection between loss of segmental material and the gain of tonal contrasts. For instance, in an early stage of the development of Vietnamese, final /h/ and /&#x0294;/ can be assumed to have been preceded by phonetically falling and rising tonal intonational contours, respectively. Subsequently final /h/ and /&#x0294;/ were both lost, and the erstwhile phonetic prosodic difference on the preceding vowels turned into a phonological, tonal distinction (<xref ref-type="bibr" rid="ref13">Haudricourt, 1954</xref>). Earlier (some time between 500BCE and 500CE), Chinese had undergone a similar development (<xref ref-type="bibr" rid="ref38">Sagart, 1999</xref>). Such developments are not restricted to SE Asia. For instance, at least four languages of Mexico and Guatemala pertaining to different branches of the Mayan family have developed contrastive tones in the context of former laryngeals (<xref ref-type="bibr" rid="ref3">Bennett, 2016</xref>, p. 497&#x2013;498). Although it is far from all cases of tonogenesis that involve a loss of segmental material (cf. <xref ref-type="bibr" rid="ref19">Michaud and Sands, 2020</xref> for a recent review), documented cases of this particular pathway justifies the interpretation of an inverse correlation between tonal contrasts and word length as being non-spurious.</p>
<p>This paper seems to be the first to investigate the relationship between word length and tones across languages. Previously the relationship between word length and segment inventory sizes was examined, with somewhat ambiguous results. <xref ref-type="bibr" rid="ref25">Nettle (1995)</xref> suggested the existence of an inverse relationship between word length and inventory sizes, but on a very small empirical basis. <xref ref-type="bibr" rid="ref22">Moran and Blasi (2014</xref>, p. 234&#x2013;236) and <xref ref-type="bibr" rid="ref46">Wichmann and Holman (2023)</xref> brought more data to the table, also finding an inverse relationship, but the latter authors were not able to confirm a statistical significance of the findings. Further afield, <xref ref-type="bibr" rid="ref16">Maddieson (2007)</xref> found positive correlations between the sizes of vowel and consonant inventories and the complexity of tonal systems, whereas syllable complexity and tone were negatively correlated according to his study.</p>
</sec>
<sec id="sec2" sec-type="materials">
<title>Materials</title>
<p>Conceivably, there are many options for obtaining information on word length and tonal distinctions across different languages. Potential sources for such information include textual corpora, dictionaries, grammars, and typological databases, where the last-mentioned type of source could possibly be constructed from any selection of the first three kinds of sources. The choices of sources of information for the present paper have been guided by two major criteria: comparability and coverage. Those criteria have led to the selection of large typological databases as sources of information. As in <xref ref-type="bibr" rid="ref46">Wichmann and Holman (2023)</xref>, the 40-item word lists in the lexical ASJP database (<xref ref-type="bibr" rid="ref48">Wichmann et al., 2020</xref>) were chosen as a source of word length data because they represent around &#x00BE; of the world&#x2019;s languages, which makes for a better coverage than any other source. Additionally, the data are comparable since the words in the list pertain to one and the same fixed set of meanings and are transcribed phonemically in a standard way. As for the information on tone system, this comes from Phoible (<xref ref-type="bibr" rid="ref23">Moran and McCloy, 2019</xref>) with some additions from the WALS chapter on tone (<xref ref-type="bibr" rid="ref17">Maddieson, 2013</xref>) and The Database of Eurasian Phonological Inventories (<xref ref-type="bibr" rid="ref27">Nikolaev, 2018</xref>). These sources together offer a coverage of around &#x00BC; of the world&#x2019;s languages and consistency in the type of data targeted, namely phonological systems. Although the description of phonemic distinctions may vary between researchers (<xref ref-type="bibr" rid="ref20">Moran, 2012</xref>), the counts of tonal distinctions are at least similar in the sense that they aim to include all distinctions attested in a given language (as opposed to, say, all distinctions attested in some corpus).</p>
<p>One criterion that might be considered in addition to coverage and comparability is representativeness. The average length of items pertaining to a short word list is not necessarily representative of the lexicon as a whole or mean word length in usage. Nor is a number of tonal distinctions necessarily representative of actual usage, since two languages might each make use of the same number of distinctions but with widely different distributions of frequencies. There are, however, two major reasons why the criterion of representativeness is not given priority here. First, representativeness is not a trivial notion, but one that requires potentially controversial assumptions concerning the entity represented. If a language is considered to be the sum of all discourses produced using a certain code, then a representative sample would be a large corpus covering different genres and modalities. If a language is considered to be a set of lexical and phonological elements combined through some syntagmatic rules, then a representative sample might be a selection of lexical elements, perhaps subjected to selected syntagmatic operations. Thus, it is not clear how to even define a criterion of representativeness. Another major reason why representativeness is not given priority is that it will often clash both with the criterion of comparability, which is a principle that cannot be relinquished, as well as with the criterion of coverage, which is more flexible than comparability, but also important. For instance, among the many corpora existing for various languages, most would not be comparable since they would be different in contents, treating different topics and representing different genres, as well as in form, being encoded in different orthographies. Moreover, for many languages no corpora are available at all, compromising the criterion of coverage.</p>
<p>The optimal sample is neither easy to define nor easy to obtain. Therefore it would be a relief to be able to show that various sources of word length data actually produce similar results. In the following I will report on some analyses indicating the degree to which this wish may be fulfilled. Briefly, I compare counts of word length based on ASJP 40-item lists with (1) 100-item lists from ASJP, (2) 985-item lists from NorthEuralex (<xref ref-type="bibr" rid="ref7">Dellert et al., 2020</xref>), and (3) corpora from TeDDi (<xref ref-type="bibr" rid="ref21">Moran et al., 2022</xref>) representing (3a) Bible texts, and (3b) versions of the Universal Declaration of Human Rights. The reader who wishes to skip the details may jump to <xref rid="tab1" ref-type="table">Table 1</xref> where the results are gathered.</p>
<table-wrap position="float" id="tab1">
<label>Table 1</label>
<caption>
<p>Correlations (Pearson&#x2019;s <italic>r</italic>) between mean word length of 40-item ASJP word lists and other data sources (in all cases <italic>p</italic>&#x2009;&#x003C;&#x2009;0.001).</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Data source</th>
<th align="center" valign="top">
<italic>r</italic>
</th>
<th align="center" valign="top">
<italic>N</italic>
</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">100-item ASJP lists</td>
<td align="center" valign="top">0.94</td>
<td align="center" valign="top">1250</td>
</tr>
<tr>
<td align="left" valign="top">985-item NorthEuraLex lists in ASJPcode</td>
<td align="center" valign="top">0.78</td>
<td align="center" valign="top">105</td>
</tr>
<tr>
<td align="left" valign="top">985-item NorthEuraLex lists in original orthography</td>
<td align="center" valign="top">0.68</td>
<td align="center" valign="top">92</td>
</tr>
<tr>
<td align="left" valign="top">Universal Declaration of Human Rights</td>
<td align="center" valign="top">0.58</td>
<td align="center" valign="top">36</td>
</tr>
<tr>
<td align="left" valign="top">Bible texts</td>
<td align="center" valign="top">0.60</td>
<td align="center" valign="top">49</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Before describing the comparisons with other sources of word length data, let me present the data actually used. For the present purposes a word is defined as the typical source of an ASJP item, which is an entry in a dictionary marked as a single, separate string by leading and trailing spaces and providing a translational equivalent of a specific concept commonly lexicalized throughout the languages of the world. Mean word length of a language is defined as the mean across such ASJP items. If two synonyms are given for a certain concept, an average length is used here, and if more than two synonyms are given, only the two first ones listed are taken into account. Phrases (anything with one or more spaces in it) are ignored. All identifiable inflectional affixes were removed during the transcription of ASJP items, so in many cases &#x2018;stem&#x2019; might actually be a more adequate description of the contents of the ASJP database, although the vast majority of the entries would be words in a normal sense. These words (or word proxies) are transcribed using ASJPcode (<xref ref-type="bibr" rid="ref5">Brown et al., 2013</xref>), a transcription system which merges phonemes into classes of phonemes but adequately represents the number of phonemes in words. It operates with 34 consonant and 7 vowels symbols, a nasalization symbol, and modifiers indicating that sequences of two or three symbols are to be interpreted as single phonemes. Additionally, there is a symbol (%) to indicate that a word is a borrowing (this is not systematically applied). For each language as defined by ISO 639-3, the word length of a certain item on the 40-item list is averaged across the word lists pertaining to one and the same ISO 639-3 language, in case more than one is available (on average there is close to two word lists per language). The following list represents the doculect <sc>english</sc>. It is not necessarily a typical list, but it is one that any reader can immediately relate to (for other examples, the reader may visit <ext-link xlink:href="https://asjp.clld.org/languages" ext-link-type="uri">https://asjp.clld.org/languages</ext-link>). The total count of phonemes in this list is 134, which, divided by the list length of 40, yields an average word length of 3.35.</p>
<p>Ei &#x2018;I,&#x2019; yu &#x2018;you,&#x2019; wi &#x2018;we,&#x2019; w3n &#x2018;one,&#x2019; tu &#x2018;two,&#x2019; %prs3n &#x2018;person,&#x2019; fiS &#x2018;fish,&#x2019; dag &#x2018;dog,&#x2019; laus &#x2018;louse,&#x2019; tri &#x2018;tree,&#x2019; lif &#x2018;leaf,&#x2019; %skin &#x2018;skin,&#x2019; bl3d &#x2018;blood,&#x2019; bon &#x2018;bone,&#x2019; horn &#x2018;horn,&#x2019; ir &#x2018;ear,&#x2019; Ei &#x2018;eye,&#x2019; noz &#x2018;nose,&#x2019; tu8 &#x2018;tooth,&#x2019; t3N &#x2018;tongue,&#x2019; ni &#x2018;knee,&#x2019; hEnd &#x2018;hand,&#x2019; brEst &#x2018;breast,&#x2019; liv3r &#x2018;liver,&#x2019; driNk &#x2018;drink,&#x2019; si &#x2018;see,&#x2019; hir &#x2018;hear,&#x2019; dEi &#x2018;die,&#x2019; k3m &#x2018;come,&#x2019; s3n &#x2018;sun,&#x2019; star &#x2018;star,&#x2019; wat3r &#x2018;water,&#x2019; ston &#x2018;stone,&#x2019; fEir &#x2018;fire,&#x2019; pE8 &#x2018;path,&#x2019; %maunt3n &#x2018;mountain,&#x2019; nEit &#x2018;night,&#x2019; ful &#x2018;full,&#x2019; nu &#x2018;new,&#x2019; nem &#x2018;name.&#x2019;</p>
<p>The word length data used in the analyses of this paper is drawn from a file called Data-01 ASJP data raw.txt, available at <ext-link xlink:href="https://zenodo.org/record/6344024" ext-link-type="uri">https://zenodo.org/record/6344024</ext-link>. The file was previously used in <xref ref-type="bibr" rid="ref46">Wichmann and Holman (2023)</xref>. It contains columns for ISO 639-3 codes, doculect names, language codes and family classifications from WALS (<xref ref-type="bibr" rid="ref9">Dryer and Haspelmath, 2013</xref>) and Glottolog (<xref ref-type="bibr" rid="ref11">Hammarstr&#x00F6;m et al., 2021</xref>), coordinates, population figures from Ethnologue (<xref ref-type="bibr" rid="ref41">Simons and Fennig, 2017</xref>), word length averaged over the 40 ASJP items and over the entire 100-item Swadesh list when available; there are also assignments of &#x2018;area,&#x2019; &#x2018;continent,&#x2019; and &#x2018;macrocontinent&#x2019; from Autotyp (<xref ref-type="bibr" rid="ref4">Bickel et al., 2017</xref>), as well as some other columns of less relevance in the present context. Word length data can be obtained from ASJP for 5289 languages (here and henceforth as defined by ISO 639-3).</p>
<p>In order to estimate the extent to which word length data based on the 40 ASJP items compares to some other sources of word length data I drew samples from the following sources: 100-item lists that are also part of the ASJP database, longer word lists in NorthEuraLex (<xref ref-type="bibr" rid="ref7">Dellert et al., 2020</xref>) and text corpora from TeDDi (<xref ref-type="bibr" rid="ref21">Moran et al., 2022</xref>). These comparanda are meant to represent samples that may be conceived of as being more representative of the involved languages than the 40 ASJP items. Mean word length for 100-item word lists are directly obtained from the same dataset used here for the 40-item lists. NorthEuraLex contains 1016-item word lists for 107 Eurasian language varieties in transcriptions that include standard orthographies and, conveniently, also ASJPcode. In order to enhance comparability I removed the least attested items (31 items attested in less than 98 languages). I also removed two languages that had been excluded from the ASJP data for not being anyone&#x2019;s current mother tongue, namely Latin and Standard Arabic. For the remaining 105 985-item word lists average word lengths were computed from the ASJPcode transcriptions. Additionally, for 92 languages associated with alphabetical writing systems, word length was computed from orthographical forms. As examples of text corpora I extracted Universal Declaration of Human Rights texts and Bible texts from TeDDi. TeDDi is conceived of as a sort of complement to WALS (<xref ref-type="bibr" rid="ref9">Dryer and Haspelmath, 2013</xref>), containing corpora for 89 languages that belong to the core WALS sample of 100 languages.<xref rid="fn0003" ref-type="fn"><sup>1</sup></xref> While the corpora are generally heterogeneous, Bible texts and Universal Declaration of Human Rights texts recur among them. Only languages represented in alphabetical writing systems could be used. Left were 36 languages with Universal Declaration of Human Rights texts and 49 languages with Bible texts from which to extract mean word lengths. Since TeDDi has a good areal and genealogical spread of languages and offers the corpora nicely organized in a single R object it is a convenient choice of sources. It goes without saying that larger sets of corpora could have been used, but for the present purposes this would seem unnecessary.</p>
<p>Results of comparing word length counts across languages for the different sources are displayed in <xref rid="tab1" ref-type="table">Table 1</xref>. When increasing the representativeness of the word lists from 40 to 100 and then to 985 items the correlation changes from 1.00 to 0.94 and then to 0.78. From the point of view of the presumably more representative sample this can be interpreted as an increase in adequacy, first by 0.06 (1.00&#x2013;0.94) when going from 40 to 100 items and then an additional 0.16 (0.94&#x2013;0.78) when going from 100 to 985 items. Continuing down the table we observe a difference of 0.10 correlation between the ASJPcode and original orthographical NorthEuraLex word lists. In this case the difference can only be interpreted as a loss, because the systematic ASJPcode should make for better comparability than traditional orthographic forms. When moving to the corpora, we observe a correlation of ~0.6. Because of the two different versions of transcriptions contained in NorthEuraLex we expect that a systematic phonemic transcription of a corpus would have yielded an around ~0.1 better correlation with the 40-item ASJP lists, i.e., the correlation with corpora would then be ~0.7.</p>
<p>As discussed above, representativeness is not a straightforward and uncontroversial notion. Still, we might consider either more extensive word lists or corpora as more representative of a language than the 40 ASJP items. Results using short word lists would be more different from results using corpora than from results using long word lists, but in either case the results would not be radically different if we were able to obtain systematic, phonemic transcriptions for the long word lists or the corpora. Such transcriptions, however, are rarely available, compounding the general lack of availability for long word lists and corpora. Thus, to conclude these experiments regarding alternative data sources: alternative data sources might be preferable from the point of view of representativeness, but for many practical purposes they would be problematical because of the challenges incurred by limitations on availability and the existence of different orthographical systems. Moreover, the relatively high correlations found between 40-item ASJP lists and the other data sources suggest that the short word lists can reasonably be used as a proxy for those other kinds of more extensive sources.</p>
<p>Data on the number of tonal distinctions can be obtained from Phoible (<xref ref-type="bibr" rid="ref23">Moran and McCloy, 2019</xref>), with a few modifications. Phoible includes data from The Database of Eurasian Phonological Inventories (<xref ref-type="bibr" rid="ref27">Nikolaev, 2018</xref>, henceforth EURPhon), but the data on tones were not included. Instead, all languages from EURPhon are represented as not having tones. Therefore, the EURPhon data in Phoible were removed and replaced by data coming directly from EURPhon. Moreover, a few errors were spotted relating to language supposedly not having tones in the Phoible &#x201C;PH&#x201D; dataset.<xref rid="fn0004" ref-type="fn"><sup>2</sup></xref> Since a &#x2018;0&#x2019; seems to sometimes means &#x2018;not applicable&#x2019; rather than absence of tones, all data points pertaining to the PH dataset encoding a language with 0 for tones were removed. Data from another 257 languages can be added from the WALS chapter on tone (<xref ref-type="bibr" rid="ref17">Maddieson, 2013</xref>), extending the data available on the simple presence or absence of tones. After excluding languages not suitable for the present research (artificial, creoles, pidgins, fake, speech registers, unclassified, mixed languages, languages for which less than 20 out of the 40 items are attested) and extracting the data overlapping between ASJP and the sources for tonal data, 1,380 languages remain. That is, for 1,380 languages both word length counts and counts of tonal distinctions are available. For an additional 108 languages there was data on presence vs. absence of tones, but not the number of tones (beyond 0). Just as for the word length counts, the unit of analysis is a language as defined by ISO 639-3. Therefore, in case more than one inventory is available for an ISO 639-3 language, the number of tones is averaged.</p>
<p>Finding good alternatives to such data on tonal distinctions coming from typological databases seems even less viable than the alternatives to word length data that we discussed. Plausibly it might be an advantage if data on tonal distinctions came directly from the same sample of words from which word length counts are produced, for instance. But many of the sources of lexical data used do not adequately record tones, and even for those that do, the ASJP database does not include this information.</p>
</sec>
<sec id="sec3" sec-type="methods">
<title>Methods</title>
<p>R scripts (<xref ref-type="bibr" rid="ref34">R Core Team, 2022</xref>) for processing the data from ASJP, Phoible, and WALS and for performing analyses is available online (see the Data Availability Statement). The relationship between tones and word length is explored in a variety of ways. A linear mixed effects model was fitted using the lme4 package (<xref ref-type="bibr" rid="ref1">Bates et al., 2015</xref>). The lme4 package is again involved in a logistic regression analysis. These analyses mainly served to generalize across language families. Various aspects of data preparation and plotting involved the dplyr (<xref ref-type="bibr" rid="ref53">Wickham et al., 2023</xref>), tibble (<xref ref-type="bibr" rid="ref24">M&#x00FC;ller and Wickham, 2022</xref>), ggplot2 (<xref ref-type="bibr" rid="ref51">Wickham, 2016</xref>), rworldmap (<xref ref-type="bibr" rid="ref42">South, 2011</xref>), and colorspace (<xref ref-type="bibr" rid="ref54">Zeileis et al., 2020</xref>) packages.</p>
<p>In order to investigate whether a negative correlation between word length and the number of tonal distinctions also shows up within families I carried out linear regression and phylogenetic correlation. The sign and magnitude of the linear regression provides information on the general nature of the relationship. Non-independence of the data, however, render <italic>p</italic>-values non-trustworthy. Instead, the phylogenetic correlation analysis (<xref ref-type="bibr" rid="ref28">Pagel, 1994</xref>, <xref ref-type="bibr" rid="ref29">1997</xref>, <xref ref-type="bibr" rid="ref30">1999</xref>) serves to estimate the likelihood of a model where the word length and the number of tonal distinctions are assumed to be correlated. This analysis required special efforts because some components of the pipeline were not available and had to be developed. The idea of the analysis is to map the word length and tone data onto phylogenetic trees having distinctive branch length in order to see whether the evolutions of the two features are coupled. In order to achieve this, I used trees from Glottolog (<xref ref-type="bibr" rid="ref11">Hammarstr&#x00F6;m et al., 2021</xref>) pruned such that only those languages appear for which lexical distances could be computed and for which data on tones and word length were available. The Glottolog trees were then supplied with branch lengths based on lexical distances from ASJP, and the phylogenetic correlation analysis could be carried out using BayesTraits (<xref ref-type="bibr" rid="ref31">Pagel et al., 2004</xref>).</p>
<p>Continuing with more detail on the pipeline for correlated evolution, the first step was to compute lexical distances in order to be able to supply branch lengths. In a formally similar kind of analysis of correlated evolution involving some linguistic traits, <xref ref-type="bibr" rid="ref40">Shcherbakova et al. (2022)</xref> used the ASJP-based global tree of <xref ref-type="bibr" rid="ref15">J&#x00E4;ger (2018)</xref> as well as a few Bayesian trees from the literature representing larger language families. The alternative of using Glottolog trees with added branch lengths ensures a degree of consensus regarding the structure of the tree as well as transparency and consistency; it avoids the awkward notion of a single world language family; and it allows for using the latest updates of ASJP (here version 20 is used; J&#x00E4;ger&#x2019;s tree is based on version 17). The lexical distances represent averages of a length-normalized Levenshtein distances (edit distances) across word pairs on the 40-item ASJP word lists: for each pair of words referring to the same concept the Levenshtein distance is found. (A convenient function for this is the adist() function of Base R). It is normalized by the length of the longest of the two strings compared. In various papers since <xref ref-type="bibr" rid="ref14">Holman et al. (2008)</xref> this has been referred to as LDN (&#x2018;Levenshtein Distance Normalized&#x2019;). <xref ref-type="bibr" rid="ref47">Wichmann et al. (2010a)</xref> showed empirically that a further modified version of the Levenshtein distance (called LDND for &#x2018;Levenshtein Distance Normalized Divided&#x2019;) is better for comparisons potentially involving unrelated languages, but since we are here only comparing related languages the less computationally intensive LDN distance suffices. It has been implemented in the interactive software of <xref ref-type="bibr" rid="ref45">Wichmann (2023)</xref>. This has many ways of selecting doculects and various choices of analyses and output. For the present purposes I exclude proto-languages, ancient attested languages, languages gone extinct between ancient times and around 1700; I choose only one doculect per ISO 639-3 language, namely the one represented by the longest word list; and I restrict word lists to those that have at least 20 items. The program operates through menus asking for input from the user. For instance, in order to produce an LDN matrix for Nilotic in an output file called Nilotic_LDN.txt the user input would supply the following 15 responses when the program is first used (using spaces to separate responses): 2 1 2 1 2 1 2 Nilotic 1238&#x2009;m 20 1 3 a 2 Nilotic_LDN.txt. For convenience, the relevant output matrices are supplied online (see Data Availability Statement).</p>
<p>Continuing with more detail on the pipeline for correlated evolution, adding lexical distances from ASJP to Glottolog trees requires a matching of ASJP doculect names and Glottocodes. This is mainly achieved using the file languages.csv from <ext-link xlink:href="https://zenodo.org/record/7079637" ext-link-type="uri">https://zenodo.org/record/7079637</ext-link>, with some modifications of matches: in cases where an ASJP doculect is matched with a glottocode representing the &#x2018;dialect&#x2019; or &#x2018;family&#x2019; level, the phylogenetically closest &#x2018;language&#x2019;-level glottocode is assigned instead. This procedure makes sense conceptually and is also required technically because later in the pipeline the keep_as_tip() function of the glottoTrees package (<xref ref-type="bibr" rid="ref37">Round, 2021</xref>) will be used for tree pruning, and this function will stop and issue an error message if the result of pruning a tree would leave a taxon as a descendant of another taxon. For instance, <sc>standard</sc>_<sc>albanian</sc> is assigned to the glottocode alba1267, which is a &#x2018;family&#x2019;-level label belonging to a higher taxonomic level than, for instance, <sc>albanian</sc>_<sc>tosk</sc> (tosk1239). In fact, the two doculects should both be assigned to tosk1239, since the Tosk dialect is the basis for the standard language. More commonly, however, the problem is that a doculect is assigned to the &#x2018;dialect&#x2019; level. For instance, <sc>bosnian</sc> is assigned to &#x2018;Bosnian standard&#x2019; (bosn1245), which itself is a &#x2018;dialect&#x2019; of &#x2018;Eastern Herzegovinian Shtokavian&#x2019; (east2821), which itself is a &#x2018;dialect&#x2019; of &#x2018;New Shtokavian&#x2019; (news1236), which itself is a &#x2018;dialect&#x2019; of &#x2018;Shtokavski&#x2019; (shto1241), which itself is a &#x2018;dialect&#x2019; of the &#x2018;Serbian-Croatian-Bosnian&#x2019; (sout1528) &#x2018;language.&#x2019; While this is the only case encountered of as many as four levels of &#x2018;dialect&#x2019; it receives the same treatment as less complicated cases, namely a direct reassignment of the dialect to the language level (in this case changing bosn1245 to sout1528).</p>
<p>After having prepared distance matrices for those ASJP languages for which information on tonal distinctions are available and having assigned glottocodes to them, the Glottolog trees are pruned so as to only contain the languages also appearing in the distance matrices. This is done using the keep_as_tip() function of glottoTrees (version 0.1; <xref ref-type="bibr" rid="ref37">Round, 2021</xref>). While this works smoothly once the problems mentioned in the previous paragraphs are taken care of, its output needs further processing in case internal non-branching nodes are retained after pruning. For instance, let two final taxa (tips) A and B be united under an internal node Int. In the Newick notion<xref rid="fn0005" ref-type="fn"><sup>3</sup></xref> such a tree would be represented as ((A,B)int,C). If B is removed during the pruning process the function will still leave Int within the tree, even if this node is not branching, in Newick notion: ((A)int,C). Such &#x2018;phantom&#x2019; nodes are not tolerated by nnls.tree() of Phangorn 2.10.0 (<xref ref-type="bibr" rid="ref39">Schliep, 2011</xref>), the function used here to supply tree with distinctive branch lengths. Indeed, they are generally not foreseen by phylogenetic software. For instance, MEGA (<xref ref-type="bibr" rid="ref43">Tamura et al., 2021</xref>) will not be able to display a tree with non-branching nodes. Fortunately, there is a simple solution to this problem. Since the internal nodes and placeholder branch lengths of 1 of the Glottolog trees are not needed, these features can be removed using regular expressions. This will leave only tip labels and brackets, easing further edits to the Newick format. A non-branching node will appear as a set of &#x2018;phantom&#x2019; brackets not containing commas not already contained in other brackets contained within the &#x2018;phantom&#x2019; brackets. In our simplest-possible example there would be a set of &#x2018;phantom&#x2019; brackets left around A as it is deprived of its sister B: ((A,B)int,C) &#x2192; ((A,B),C) &#x2192; ((A),C). In cases where the pruned taxon is not the terminal sister of a single other taxon, some further look-around is required to find the two friends making up a pair of &#x2018;phantom&#x2019; brackets, as in the case of (((A,B)),(C,D)) &#x2190; (((A,B),E),(C,D)), where the culprits are the extra brackets around (A,B). These cases will be identifiable as two consecutive opening brackets that are members of a set of brackets which includes closing brackets which are likewise consecutive. Based on these insights, an algorithm was implemented in my fix.non.br() function in the phylogenetic_correlation.R script supplied along with this paper. Other functions, from various packages, that were used in the tree manipulation procedures included read.tree(), write.tree(), drop.tip(), and write.nexus() from ape 5.7 (<xref ref-type="bibr" rid="ref32">Paradis and Schliep, 2019</xref>); and str_split() and str_sub() from stringr 1.5.0 (<xref ref-type="bibr" rid="ref52">Wickham, 2022</xref>).</p>
<p>At this point in the pipeline a distance matrix and a Newick tree is available for each language family (where a family is required to have 6 or more members). This is the input needed for Phangorn&#x2019;s nnls.tree() function, which is used for supplying the Glottolog trees with branch lengths. Previously <xref ref-type="bibr" rid="ref6">Dediu (2018)</xref> similarly used this function to supply branch lengths from various sources to language family trees of different extractions (Ethnologue, WALS, Autotyp, Glottolog), and I am inspired by this work but use my own implementation of the process. What nnls.tree() does, summarily stated, is to estimate branch lengths such that patristic distances among taxa, i.e., the distances between taxa along the tree, best approximate the distances in the supplied matrix. This is done by applying the least squares criterion, minimizing the sum of squared errors. A blog post by <xref ref-type="bibr" rid="ref36">Revell (2011)</xref> provides an entry point for better comprehension. It is of interest to look at how well the resulting patristic distances fit the original LDN distances. This is done for each family using the mantel.rtest() of ade4 1.7.19 (<xref ref-type="bibr" rid="ref8">Dray and Dufour, 2007</xref>). The resulting <italic>r</italic> values, which are all significant at the <italic>p</italic>&#x2009;&#x003C;&#x2009;0.01 level, are reported in <xref rid="tab2" ref-type="table">Table 2</xref>, in descending order. I am not aware of similar tests of other, comparable branch length fitting outcomes, so it is difficult to know what to require from the results, but the fits certainly seem good enough to at least pass a sanity test: the results are approximately normally distributed around a high mean of 0.93.</p>
<table-wrap position="float" id="tab2">
<label>Table 2</label>
<caption>
<p>Results of mantel tests for LDN and patristic distances in trees supplied with branch lengths.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Family</th>
<th align="center" valign="top">
<italic>r</italic>
</th>
<th align="left" valign="top">Family</th>
<th align="center" valign="top">
<italic>r</italic>
</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">Otomanguean</td>
<td align="center" valign="top">0.981</td>
<td align="left" valign="top">Austronesian</td>
<td align="center" valign="top">0.938</td>
</tr>
<tr>
<td align="left" valign="top">Central Sudanic</td>
<td align="center" valign="top">0.972</td>
<td align="left" valign="top">Nuclear Trans New Guinea</td>
<td align="center" valign="top">0.928</td>
</tr>
<tr>
<td align="left" valign="top">Tai-Kadai</td>
<td align="center" valign="top">0.970</td>
<td align="left" valign="top">Indo-European</td>
<td align="center" valign="top">0.928</td>
</tr>
<tr>
<td align="left" valign="top">Mande</td>
<td align="center" valign="top">0.967</td>
<td align="left" valign="top">Afro-Asiatic</td>
<td align="center" valign="top">0.928</td>
</tr>
<tr>
<td align="left" valign="top">Kadugli-Krongo</td>
<td align="center" valign="top">0.960</td>
<td align="left" valign="top">Sino-Tibetan</td>
<td align="center" valign="top">0.923</td>
</tr>
<tr>
<td align="left" valign="top">Nilotic</td>
<td align="center" valign="top">0.955</td>
<td align="left" valign="top">Atlantic-Congo</td>
<td align="center" valign="top">0.899</td>
</tr>
<tr>
<td align="left" valign="top">Athabaskan-Eyak-Tlingit</td>
<td align="center" valign="top">0.944</td>
<td align="left" valign="top">Ta-Ne-Omotic</td>
<td align="center" valign="top">0.810</td>
</tr>
<tr>
<td align="left" valign="top">Austroasiatic</td>
<td align="center" valign="top">0.939</td>
<td align="left" valign="top">Salishan</td>
<td align="center" valign="top">0.772</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>As the last element of the correlated evolution pipeline the software BayesTraits in its most recent instantiation, version 4.0.1 (<xref ref-type="bibr" rid="ref18">Meade and Pagel, 2023</xref>), is put to work. Similarly to <xref ref-type="bibr" rid="ref40">Shcherbakova et al. (2022)</xref>, I follow the recommendations of the BayesTraits manual for testing correlations between continuous traits (<xref ref-type="bibr" rid="ref18">Meade and Pagel, 2023</xref>, p. 37&#x2013;38). The assumption here is that traits evolve as random walks. To estimate whether two traits are coevolving, a complex model assuming a correlation is compared with a simple model in which the correlation is set to zero. The strength of the complex model over the simple one is estimated through a log Bayes Factor, calculated as 2 &#x002A; (log marginal likelihood complex model &#x2013; log marginal likelihood simple model). These log Bayes Factors may be interpreted as in <xref rid="tab3" ref-type="table">Table 3</xref>, following <xref ref-type="bibr" rid="ref35">Raftery (1996)</xref>.</p>
<table-wrap position="float" id="tab3">
<label>Table 3</label>
<caption>
<p>Interpretations of log Bayes Factors (from <xref ref-type="bibr" rid="ref35">Raftery, 1996</xref>, p. 165).</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Log Bayes Factors</th>
<th align="left" valign="top">Evidence for alternative hypothesis</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">&#x003C;0</td>
<td align="left" valign="top">Negative (supports null hypothesis)</td>
</tr>
<tr>
<td align="left" valign="top">0&#x2013;2</td>
<td align="left" valign="top">Barely worth mentioning</td>
</tr>
<tr>
<td align="left" valign="top">2&#x2013;5</td>
<td align="left" valign="top">Positive</td>
</tr>
<tr>
<td align="left" valign="top">5&#x2013;10</td>
<td align="left" valign="top">Strong</td>
</tr>
<tr>
<td align="left" valign="top">&#x003E;10</td>
<td align="left" valign="top">Very strong</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
<sec id="sec4" sec-type="results">
<title>Results</title>
<p>We begin to explore the nature of the relationship between word length and the number of tonal distinctions by means of the boxplots in <xref rid="fig1" ref-type="fig">Figure 1</xref>. Each boxplot represents mean word length values for a certain number of tonal distinctions. Sometimes, when more than one language variety is involved, the number of tonal distinctions of an ISO 639-3 language (the unit of analysis) is not a whole number. For the purpose of the graph, the number has then been rounded off to the nearest integer. Small squares represent means. The fitted line is not based on any kind of binning but represents the linear fit of all values of number of tonal distinctions and mean word length. Although this fit over the entire range is decent (<italic>R</italic><sup>2</sup>&#x2009;=&#x2009;0.196), the graph suggests that the correlation mainly holds for values of tonal distinctions from 0 to 3, while the relationship for values in the range 4&#x2013;10 is at best weak. Apparently there is a lower limit on vowel length of 2&#x2013;3 segments that languages cannot cross without losing too much in terms of expressive means. But once this limit is reached, tonal systems can still develop in complexity for reasons other than through compensation for segment loss. Referring to three or more tones as &#x2018;several,&#x2019; we can say that mean word length is a strong predictor of whether a language will have zero, one, two or several tones. The number of tones above three, however, would seem not to depend appreciably on this factor, at least as far as we can judge from the available data, which is relatively limited for the complex systems. Still, in order to avoid manufacturing of results, we do not combine three or more tones in one bin, but continue to operate with the original range of values in subsequent analyses.</p>
<fig position="float" id="fig1">
<label>Figure 1</label>
<caption>
<p>Boxplots of mean word length for different numbers of tonal distinctions. Small black squares represent means and the dashed line is a linear fit of all raw values of mean word length and the number of tonal distinctions.</p>
</caption>
<graphic xlink:href="fpsyg-14-1128461-g001.tif"/>
</fig>
<p>Before exploring the relationship between the number of tonal contrasts and mean word length further, we ask whether the relationship is statistically significant in the first place. The question is answered by formulating a linear mixed effects model with the number of tonal contrasts as a function of mean word length (predictor variable) and random effects represented by Glottolog family membership and membership of one of the following &#x2018;continents&#x2019; of Autotyp: Africa, Western and Southwestern Eurasia, North-Central Asia, South and Southeast Asia, New Guinea and Oceania, Australia, Eastern North America, Western North America, Central America, and South America (when a family is spread over more than one continent all members are assigned to just one continent, namely the one from which scholars would normally assume the family to have originated, cf. discussion of received views in <xref ref-type="bibr" rid="ref49">Wichmann et al., 2010b</xref>; a list of the decisions taken is in the script tones.R, provided online). When trying to estimate both slopes and intercepts for the random effects singular fits arose, so here only the intercepts are estimated. The summary of the model is found in <xref rid="box1" ref-type="boxed-text">Box 1</xref>.</p>
<boxed-text id="box1" position="float">
<sec id="sec800">
<label>BOX 1.</label>
<title>Summary of linear mixed effects model with number of tonal contrasts as a function of mean word length (predictor) and family &#x0026; continent (random effects).</title>
<p>&#x2003;Linear mixed model fit by maximum likelihood [&#x2018;lmerMod&#x2019;]</p>
<p>&#x2003;Formula: count_tones ~ forty_mean + (1 | continent) + (1 | glot_fam)</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;Data: pho2</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;AIC&#x2003;&#x2003;&#x2003;&#x2003;BIC&#x2003;&#x2003;&#x2003;logLik&#x2003;&#x2003;&#x2003;deviance&#x2003;&#x2003;&#x2003;&#x2003;df.resid</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;4781.9&#x2003;&#x2003;4808.0&#x2003;&#x2003;-2385.9&#x2003;&#x2003;&#x2003;&#x2003;4771.9&#x2003;&#x2003;&#x2003;&#x2003;1375</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;Scaled residuals:</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;Min&#x2003;&#x2003;&#x2003;&#x2003;1Q&#x2003;&#x2003;&#x2003;Median&#x2003;&#x2003;&#x2003;3Q&#x2003;&#x2003;&#x2003;Max</p>
<p>&#x2003;&#x2003;&#x2003;-2.4076&#x2003;&#x2003;&#x2003;-0.3701&#x2003;&#x2003;-0.0644&#x2003;&#x2003;0.2672&#x2003;&#x2003;5.8136</p>
<p>&#x2003;Random effects:</p>
<p>&#x2003;&#x2003;&#x2003;Groups&#x2003;&#x2003;&#x2003;Name&#x2003;&#x2003;&#x2003;Variance&#x2003;&#x2003;Std.Dev.</p>
<p>&#x2003;&#x2003;&#x2003;glot_fam&#x2003;&#x2003;(Intercept)&#x2003;&#x2003;0.4780&#x2003;&#x2003;0.6914</p>
<p>&#x2003;&#x2003;&#x2003;continent&#x2003;&#x2003;(Intercept)&#x2003;&#x2003;0.2387&#x2003;&#x2003;0.4885</p>
<p>&#x2003;&#x2003;&#x2003;Residual&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;1.6970&#x2003;&#x2003;1.3027</p>
<p>&#x2003;Number&#x2003;of&#x2003;obs:&#x2003;1380,&#x2003;groups:&#x2003;glot_fam,&#x2003;178;&#x2003;continent,&#x2003;10</p>
<p>&#x2003;Fixed effects:</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;Estimate&#x2003;Std.&#x2003;&#x2003;Error&#x2003;&#x2003;t value</p>
<p>&#x2003;(Intercept)&#x00A0;&#x2003;&#x2003;&#x2003;3.12390&#x2003;&#x2003;&#x2003;&#x2003;0.34192&#x2003;&#x2003;9.136</p>
<p>&#x2003;forty_mean&#x2003;&#x00A0;&#x2003;&#x02212;0.56280&#x2003;&#x2003;&#x2003;&#x2003;0.06764&#x00A0;&#x00A0;&#x2003;&#x02212;8.320</p>
</sec>
</boxed-text>
<p>Of perhaps most interest in this output is the coefficient &#x2212;0.563, which shows that around half a tonal distinction is gained per one segment decrease of word length.</p>
<p>Using the anova() function, the full model as fitted by lmer() is compared to a reduced model where the number of tonal distinctions is a function of its own mean, with the random effects retained. The output of this comparison shows the difference between the models to be highly significant (<italic>&#x03A7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;66.32, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.0001), and the smaller AIC and BIC values and higher log likelihood of the full model also indicate the importance of mean word length as a predictor of the number of tonal contrasts (<xref rid="box2" ref-type="boxed-text">Box 2</xref>).</p>
<boxed-text id="box2" position="float">
<sec id="sec801">
<label>BOX 2.</label>
<title>Summary of comparison of full model (cf. <xref rid="box1" ref-type="boxed-text">Box 1</xref>) with a reduced model where the number of tonal contrasts is removed as predictor variable.</title>
<p>&#x2003;reduced_model: count_tones ~ 1 + (1 | continent) + (1 | glot_fam)</p>
<p>&#x2003;full_model: count_tones ~ forty_mean + (1 | continent) + (1 | glot_fam)</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;npar&#x2003;&#x2003;&#x2003;AIC&#x2003;&#x2003;&#x2003;BIC&#x2003;&#x2003;&#x2003;logLik&#x2003;deviance&#x2003;&#x2003;Chisq&#x2003;&#x2003;Df&#x2003;&#x2003;Pr(&#x0003E;Chisq)</p>
<p>&#x2003;reduced_model&#x2003;&#x2003;&#x2003;&#x2003;4&#x2003;&#x2003;4846.2&#x2003;&#x2003;4867.1&#x2003;&#x2003;-2419.1&#x2003;&#x2003;4838.2</p>
<p>&#x2003;full_model&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;5&#x2003;&#x2003;4781.9&#x2003;&#x2003;4808.0&#x2003;&#x2003;-2385.9&#x2003;&#x2003;4771.9&#x2003;&#x2003;66.315&#x2003;&#x2003;1&#x2003;&#x2003;3.843e-16&#x2003;&#x2003;&#x0002A;&#x0002A;&#x0002A;</p>
</sec>
</boxed-text>
<p><xref rid="fig2" ref-type="fig">Figures 2</xref>, <xref rid="fig3" ref-type="fig">3</xref> plot the data for, respectively, families with six or more members and continents. Black lines show the linear regressions produced by the mixed model, where only intercepts are varied. Red lines show linear regressions based on the data for individual families or continents. Typically there is a relatively good agreement between the fits of the general linear model and individual linear models for areas and larger families where tonal languages abound, while poorer fits emerge for areas and families where tonal languages are uncommon or absent; for small families some fits are probably in disagreement mainly because of small sample sizes. For continents there are similarly good agreements whenever tonal languages are common.</p>
<fig position="float" id="fig2">
<label>Figure 2</label>
<caption>
<p>Scatterplots of tonal distinctions as a function of mean word length in families with six or more members. Black lines show fits to a general mixed linear model, with intercepts varied; red lines show fits to individual linear models.</p>
</caption>
<graphic xlink:href="fpsyg-14-1128461-g002.tif"/>
</fig>
<fig position="float" id="fig3">
<label>Figure 3</label>
<caption>
<p>Scatterplots of tonal distinctions as a function of mean word length in continents. Black lines show fits to a general mixed linear model, with intercepts varied; red lines show fits to individual linear models.</p>
</caption>
<graphic xlink:href="fpsyg-14-1128461-g003.tif"/>
</fig>
<p>The family scatterplots with regression lines that tend to show negative slopes in <xref rid="fig2" ref-type="fig">Figure 2</xref> strongly suggest that once tones are more than sporadically present in a family they will have developed in tandem with decreased word length. Fitting a linear model, however, ignores the diachronic perspective&#x2014;it treats the languages as a pile of fallen leaves having no identifiable connection to specific branches in the tree that they come from. This represents a huge loss of information. In order to estimate the likelihood of a model where the developments of word length and tones are coupled, we need to include the tree structure connecting the languages in the analysis, making use of comparative methods from biology (<xref ref-type="bibr" rid="ref12">Harvey and Pagel, 1991</xref>). Specifically, we use tree topologies from Glottolog (<xref ref-type="bibr" rid="ref11">Hammarstr&#x00F6;m et al., 2021</xref>), pruned such as to contain only the languages of interest and supplied with distinctive branch lengths based on lexical distances (normalized Levenshtein distances or LDN) calculated from ASJP data (<xref ref-type="bibr" rid="ref70">Wichmann et al., 2022</xref>). Subsequently we feed the trees and the data on word length and tonal distinctions to BayesTraits (<xref ref-type="bibr" rid="ref18">Meade and Pagel, 2023</xref>). The results, again reporting on families with six or more members, are in <xref rid="tab4" ref-type="table">Table 4</xref>. This shows the log Bayes Factors, which express the amount of support for a model of correlated evolution and which may be interpreted following the guidelines in <xref rid="tab3" ref-type="table">Table 3</xref>. <xref rid="tab4" ref-type="table">Table 4</xref> also shows Pearson&#x2019;s <italic>r</italic> for the (non-phylogenetic) correlations between tones and word length (cf. the red fitted lines in <xref rid="fig2" ref-type="fig">Figure 2</xref>), mainly in order to remind us of the sign of the correlation.</p>
<table-wrap position="float" id="tab4">
<label>Table 4</label>
<caption>
<p>Log Bayes factors for phylogenetic correlation of tone and word length, Pearson&#x2019;s <italic>r</italic> for conventional correlations of the same variables, and the number of languages.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left" valign="top">Family</th>
<th align="center" valign="top">LogBF</th>
<th align="center" valign="top">
<italic>r</italic>
</th>
<th align="center" valign="top">
<italic>N</italic>
</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">Austroasiatic</td>
<td align="center" valign="top">9.43</td>
<td align="center" valign="top">&#x2212;0.525</td>
<td align="center" valign="top">32</td>
</tr>
<tr>
<td align="left" valign="top">Atlantic-Congo</td>
<td align="center" valign="top">8.13</td>
<td align="center" valign="top">&#x2212;0.297</td>
<td align="center" valign="top">337</td>
</tr>
<tr>
<td align="left" valign="top">Afro-Asiatic</td>
<td align="center" valign="top">5.93</td>
<td align="center" valign="top">&#x2212;0.196</td>
<td align="center" valign="top">96</td>
</tr>
<tr>
<td align="left" valign="top">Ta-Ne-Omotic</td>
<td align="center" valign="top">5.69</td>
<td align="center" valign="top">&#x2212;0.774</td>
<td align="center" valign="top">8</td>
</tr>
<tr>
<td align="left" valign="top">Indo-European</td>
<td align="center" valign="top">4.69</td>
<td align="center" valign="top">&#x2212;0.277</td>
<td align="center" valign="top">134</td>
</tr>
<tr>
<td align="left" valign="top">Nuclear Trans New Guinea</td>
<td align="center" valign="top">4.03</td>
<td align="center" valign="top">0.594</td>
<td align="center" valign="top">12</td>
</tr>
<tr>
<td align="left" valign="top">Central Sudanic</td>
<td align="center" valign="top">2.79</td>
<td align="center" valign="top">&#x2212;0.557</td>
<td align="center" valign="top">19</td>
</tr>
<tr>
<td align="left" valign="top">Athabaskan-Eyak-Tlingit</td>
<td align="center" valign="top">2.22</td>
<td align="center" valign="top">&#x2212;0.526</td>
<td align="center" valign="top">9</td>
</tr>
<tr>
<td align="left" valign="top">Otomanguean</td>
<td align="center" valign="top">2.03</td>
<td align="center" valign="top">&#x2212;0.444</td>
<td align="center" valign="top">9</td>
</tr>
<tr>
<td align="left" valign="top">Salishan</td>
<td align="center" valign="top">0.78</td>
<td align="center" valign="top">0.197</td>
<td align="center" valign="top">7</td>
</tr>
<tr>
<td align="left" valign="top">Sino-Tibetan</td>
<td align="center" valign="top">0.76</td>
<td align="center" valign="top">&#x2212;0.166</td>
<td align="center" valign="top">76</td>
</tr>
<tr>
<td align="left" valign="top">Mande</td>
<td align="center" valign="top">0.67</td>
<td align="center" valign="top">&#x2212;0.399</td>
<td align="center" valign="top">38</td>
</tr>
<tr>
<td align="left" valign="top">Austronesian</td>
<td align="center" valign="top">0.54</td>
<td align="center" valign="top">&#x2212;0.099</td>
<td align="center" valign="top">74</td>
</tr>
<tr>
<td align="left" valign="top">Nilotic</td>
<td align="center" valign="top">0.53</td>
<td align="center" valign="top">0.004</td>
<td align="center" valign="top">21</td>
</tr>
<tr>
<td align="left" valign="top">Kadugli-Krongo</td>
<td align="center" valign="top">&#x2212;0.24</td>
<td align="center" valign="top">0.432</td>
<td align="center" valign="top">6</td>
</tr>
<tr>
<td align="left" valign="top">Tai-Kadai</td>
<td align="center" valign="top">&#x2212;0.49</td>
<td align="center" valign="top">&#x2212;0.218</td>
<td align="center" valign="top">12</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>What emerges from <xref rid="tab4" ref-type="table">Table 4</xref> is that correlated evolution of tone and word length is supported to various degrees (LogBF &#x003E;2) in 9 cases. Another 5 cases are &#x2018;not worth talking about&#x2019; and only 2 cases (Kadugli-Krongo, Tai-Kadai) support the null hypothesis. The conventional correlation analysis indicates a negative relationship in 12 cases and a positive relationship in 4 cases. Among the latter cases, however, only Nuclear Trans New Guinea (nTNG) finds support from the phylogenetic correlation. When looking more closely at the data it turns out that only 4 out of the 12 nTNG languages are tonal. Moreover, nTNG is a contested family (<xref ref-type="bibr" rid="ref44">Wichmann, 2013</xref>). If tones are only attested in a few languages and if the genealogical relationships are uncertain we have reasons to discount these results. The Tai-Kadai languages all have 2.5&#x2013;6 tones and word lengths of 2.83&#x2013;3.35. Thus, they belong to the range of the distribution of word length and tone where the relationship breaks down, presumably because a floor on the word length has been reached (cf. <xref rid="fig1" ref-type="fig">Figure 1</xref>).</p>
<p>Another way of assessing the importance of mean word length for tones is to look at the mere presence vs. absence of tones and infer the probability of having tones as a function of mean word length. We perform this analysis using the glmer() function of the lme4 package. Presence/absence, represented by the digits 1 and 0, is fitted to the same model as earlier, with mean word length as predictor and continent and area as random effects (formulaically: <monospace>p_a ~ forty_mean + (1 | continent + (1 | glot_fam), data = pho3, family = binomial</monospace>). The summary of the model is found in <xref rid="box3" ref-type="boxed-text">Box 3</xref>.</p>
<boxed-text id="box3" position="float">
<sec id="sec807">
<label>BOX 3.</label>
<title>Summary of generalized linear mixed model with presence/absence of tone as a function of mean word length (predictor) and family &#x0026; continent (random effects).</title>
<p>&#x2003;&#x2003;&#x2003;Generalized linear mixed model fit by maximum likelihood (Laplace Approximation) [&#x2018;glmerMod&#x2019;]</p>
<p>&#x2003;Family: binomial&#x2003;( logit )</p>
<p>&#x2003;&#x2003;&#x2003;Formula: p_a ~ forty_mean + (1 | continent) + (1 | glot_fam)</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;Data: pho2</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;AIC&#x2003;&#x2003;BIC&#x2003;&#x2003;logLik&#x2003;&#x2003;deviance&#x2003;&#x2003;df.resid</p>
<p>&#x2003;&#x2003;&#x2003;1209.8&#x2003;&#x2003;1231.0&#x2003;&#x2003;&#x02212;600.9&#x2003;&#x2003;1201.8&#x2003;&#x2003;&#x2003;1484</p>
<p>&#x2003;&#x2003;&#x2003;Scaled residuals:</p>
<p>&#x2003;&#x2003;&#x2003;&#x00A0;Min&#x2003;&#x2003;1Q&#x2003;&#x2003;Median&#x2003;&#x2003;3Q&#x2003;&#x2003;Max</p>
<p>&#x2003;-3.8615&#x2003;&#x2003;-0.3149&#x2003;&#x2003;&#x02212;0.0719&#x2003;&#x2003;0.4266&#x2003;&#x2003;8.6409</p>
<p>&#x2003;&#x2003;&#x2003;Random effects: Groups&#x2003;&#x2003;Name&#x2003;&#x2003;Variance Std.Dev.</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;glot_fam (Intercept)&#x2003;&#x2003;&#x2003;2.20&#x2003;&#x2003;1.483</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;continent (Intercept)&#x2003;&#x2003;&#x2003;2.34&#x2003;&#x2003;1.530</p>
<p>&#x2003;&#x2003;&#x2003;Number of obs:&#x2003;1488,&#x2003;groups:&#x2003;glot_fam,&#x2003;201;&#x2003;continent,&#x2003;10</p>
<p>&#x2003;&#x2003;&#x2003;Fixed effects:&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;Estimate&#x2003;&#x2003;Std.&#x2003;Error z&#x2003;value&#x2003;Pr(&#x0003E;|z|)</p>
<p>&#x2003;&#x2003;&#x2003;(Intercept)&#x2003;&#x2003;3.0598642&#x2003;0.0007789&#x2003;3929&#x2003;&#x0003C;2e-16&#x2003;&#x0002A;&#x0002A;&#x0002A;</p>
<p>&#x2003;&#x2003;&#x2003;forty_mean&#x2003;&#x02212;1.1157857&#x2003;0.0007796&#x2003;&#x02212;1431&#x2003;&#x0003C;2e-16&#x2003;&#x0002A;&#x0002A;&#x0002A;</p>
</sec>
</boxed-text>
<p>Just as done for the model with the count_tones predictor, the full model with the p_a (presence/absence) predictor is compared to its counterpart without this predictor through anova(). Again we find strong support (<italic>&#x03A7;</italic><sup>2</sup>(1)&#x2009;=&#x2009;50.49, <italic>p</italic>&#x2009;&#x003C;&#x2009;0.0001, smaller AIC and BIC, higher log likelihood) for the full model (<xref rid="box4" ref-type="boxed-text">Box 4</xref>).</p>
<boxed-text id="box4" position="float">
<sec id="sec805">
<label>BOX 4.</label>
<title>Summary of comparison of full model (cf. <xref rid="box3" ref-type="boxed-text">Box 3</xref>) with a reduced model where presence/absence of tone is removed as predictor variable.</title>
<p>&#x2003;&#x2003;&#x2003;reduced_binary_model: p_a ~ 1 + (1 | continent) + (1 | glot_fam)</p>
<p>&#x2003;&#x2003;&#x2003;binary_model: p_a ~ forty_mean + (1 | continent) + (1 | glot_fam)</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;npar&#x2003;&#x2003;&#x2003;AIC&#x2003;&#x2003;&#x2003;BIC&#x2003;&#x2003;&#x2003;logLik&#x2003;&#x2003;&#x2003;deviance&#x2003;&#x00A0;Chisq&#x2003;&#x2003;Df&#x2003;&#x2003;Pr(&#x0003E;Chisq)</p>
<p>&#x2003;&#x2003;&#x2003;reduced_binary_model&#x2003;&#x2003;&#x2003;3&#x2003;&#x2003;1258.3&#x2003;&#x2003;1274.2&#x2003;&#x2003;-626.15&#x2003;&#x2003;1252.3</p>
<p>&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;&#x2003;binary_model&#x2003;&#x2003;&#x00A0;&#x2003;4&#x2003;&#x2003;1209.8&#x2003;&#x2003;1231.0&#x2003;&#x2003;-600.90&#x2003;&#x2003;1201.8&#x2003;&#x2003;50.491&#x2003;&#x2003;1&#x2003;&#x2003;1.197e-12&#x2003;&#x0002A;&#x0002A;&#x0002A;</p>
</sec>
</boxed-text>
<p>The intercept and slope are now retrieved from the summary of the model and we can infer probabilities for different values of mean word length using the plogis() function of base R&#x2019;s stats component. Results are shown in <xref rid="fig4" ref-type="fig">Figure 4</xref>. Here the curve is overlaid on a density plot of raw word length data in all the 5044 languages from ASJP available for this study. <xref rid="fig4" ref-type="fig">Figure 4</xref> shows that the probability of having tones decreases as mean word length increases from the minimum (1.93 segments) to the maximum (7.73 segments).</p>
<fig position="float" id="fig4">
<label>Figure 4</label>
<caption>
<p>Probability of having tone as a function of mean word length, as inferred through logistic regression (solid curve) overlaid on a density plot of mean word length distribution across 5044 languages in ASJP (dotted curve) and showing the overall mean of mean word lengths (red vertical line).</p>
</caption>
<graphic xlink:href="fpsyg-14-1128461-g004.tif"/>
</fig>
<p>As is well known from other surveys, including the WALS chapter on tones by <xref ref-type="bibr" rid="ref17">Maddieson (2013)</xref>, the main concentrations of tonal languages are in Subsaharan Africa and SE Asia. <xref rid="fig5" ref-type="fig">Figure 5</xref> adds information on word length to the information on the presence of tonal languages. For the purposes of this map the tonal languages in our dataset were divided into three categories according to the quartiles of mean word length to which they belong: languages with short words (1st quartile, colored blue), languages with long words (4th quartile, colored red), and languages with intermediate word length values (2nd and 3rd quartiles, colored yellow). The map reveals that associations between tones and long words tend to be proportionally more common outside of the core tonal areas (Subsaharan Africa, SE Asia) than inside them. Most strikingly, in South America and New Guinea nearly all cases of tonal languages have long or intermediately long words.</p>
<fig position="float" id="fig5">
<label>Figure 5</label>
<caption>
<p>A map of tonal language with short (blue), intermediate (yellow), and long words (red).</p>
</caption>
<graphic xlink:href="fpsyg-14-1128461-g005.tif"/>
</fig>
</sec>
<sec id="sec5" sec-type="discussions">
<title>Discussion</title>
<p>This paper has demonstrated the existence of a relationship between the number of tonal distinctions and mean word length. When controlling for membership in different world areas and language families, this relationship remains highly significant. The finding from linear mixed effect modeling that around half a tonal distinction is gained per one segment decrease of word length suggests that the relationship, apart from being significant, is also relatively strong. We did note, however, that the prediction from word length seems to break down beyond three tonal distinctions&#x2014;the number of tones that a complex system reckons with may largely be unrelated to mean word length, presumably because the limit to how short words can be on average (2&#x2013;3 segments) is reached before the limit to how many tonal distinctions a language can develop. An example of a language where tonal contrasts initially developed through segmental loss and subsequently through other means is Vietnamese. According to <xref ref-type="bibr" rid="ref13">Haudricourt (1954)</xref> a system of three tones, originally developed through segmental loss, further developed into a system of six tones through a merger of initial voiced and voiceless consonants. In general, developments of complex tone systems through the loss of a voicing distinction are common (e.g., <xref ref-type="bibr" rid="ref33">Pittayaporn and Kirby, 2017</xref> on the Tai dialect of Cao Bang and <xref ref-type="bibr" rid="ref10">Ferlus, 2009</xref> on Chinese with further references and general discussion).</p>
<p>The phylogenetic correlation analysis confirmed the existence of coupled evolution of word length and tone in many language families pertaining to the following major world macroareas: Eurasia (Austroasiatic, Indo-European), Africa (Atlantic-Congo, Afro-Asiatic, Ta-Ne-Omotic, Central Sudanic), and America (Athabaskan-Eyak-Tlingit, Otomanguean). It would be a great oversimplification to only attempt to explain the evolution of tonal systems through the loss of segmental material, though. This is not the only pathway to tones (cf. examples given in the previous paragraph and <xref ref-type="bibr" rid="ref19">Michaud and Sands, 2020</xref> for a recent overview). Moreover, it is also possible to imagine that the introduction of a tonal system could precede a loss of segments. Still, the relationship identified makes good sense in the light of a causal mechanism where a frequent initial motivation for the presence of tones would be to compensate for the lack of expressive materials as lexical morphemes become shorter. Earlier work (<xref ref-type="bibr" rid="ref50">Wichmann et al., 2011</xref>; <xref ref-type="bibr" rid="ref46">Wichmann and Holman, 2023</xref>) has demonstrated a negative correlation between word length and (log) population sizes. Taken together, the findings suggest a causal chain where larger populations lead to shorter words through general complexity reduction, and tonal systems subsequently emerge and spread among languages in order to maintain lexical distinctions, compensating for the loss of expressive means.</p>
<p>Mapping the geographical distribution of tonal languages with short vs. intermediate vs. long words suggests that the causal relationship is most prominent in Subsaharan Africa and SE Asia, two areas associated with Neolithic revolutions and large prehistorical population booms (<xref ref-type="bibr" rid="ref2">Bellwood, 2004</xref>). Thus, short words and tone tends to be an areally concentrated &#x2018;package&#x2019; which is furthermore often associated with large populations probably ultimately related to the impact of agriculture. This suggests that it would have been much less frequent among the world&#x2019;s languages in pre-Neolithic times than nowadays. Exploring the implications of the relationship between word length and tones for the prehistory of languages and their speakers requires more work and is a fascinating item for future research.</p>
</sec>
<sec id="sec6" sec-type="data-availability">
<title>Data availability statement</title>
<p>The datasets presented in this study can be found in online repositories. The names of the repository/repositories and accession number(s) can be found at: <ext-link xlink:href="https://github.com/Sokiwi/Tone-WordLength" ext-link-type="uri">https://github.com/Sokiwi/Tone-WordLength</ext-link>, <ext-link xlink:href="https://zenodo.org/record/6344024" ext-link-type="uri">https://zenodo.org/record/6344024</ext-link>.</p>
</sec>
<sec id="sec7">
<title>Author contributions</title>
<p>SW conceived and designed the study, prepared the data, performed the statistical analysis, and wrote the manuscript.</p>
</sec>
<sec id="sec8" sec-type="funding-information">
<title>Funding</title>
<p>This work was supported by the Deutsche Forschungsgemeinschaft (German Research Foundation) under Germany&#x2019;s Excellence Strategy (grant EXC 2150 390870439).</p>
</sec>
<sec id="conf1" sec-type="COI-statement">
<title>Conflict of interest</title>
<p>The author declares that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.</p>
</sec>
<sec id="sec100" sec-type="disclaimer">
<title>Publisher&#x2019;s note</title>
<p>All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.</p>
</sec>
</body>
<back>
<ref-list>
<title>References</title>
<ref id="ref1"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bates</surname> <given-names>D.</given-names></name> <name><surname>Maechler</surname> <given-names>M.</given-names></name> <name><surname>Bolker</surname> <given-names>B.</given-names></name> <name><surname>Walker</surname> <given-names>S.</given-names></name></person-group> (<year>2015</year>). <article-title>Fitting linear mixed-effects models using lme4</article-title>. <source>J. Stat. Softw.</source> <volume>67</volume>, <fpage>1</fpage>&#x2013;<lpage>48</lpage>. doi: <pub-id pub-id-type="doi">10.18637/jss.v067.i01</pub-id></citation></ref>
<ref id="ref2"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Bellwood</surname> <given-names>P.</given-names></name></person-group> (<year>2004</year>). <source>First farmers: the origins of agricultural societies</source>. <publisher-loc>Malden, MA</publisher-loc>: <publisher-name>Blackwell Publishing</publisher-name>.</citation></ref>
<ref id="ref3"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Bennett</surname> <given-names>R.</given-names></name></person-group> (<year>2016</year>). <article-title>Mayan phonology</article-title>. <source>Lang. Linguist. Compass</source> <volume>10</volume>, <fpage>469</fpage>&#x2013;<lpage>514</lpage>. doi: <pub-id pub-id-type="doi">10.1111/lnc3.12148</pub-id></citation></ref>
<ref id="ref4"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Bickel</surname> <given-names>B.</given-names></name> <name><surname>Nichols</surname> <given-names>J.</given-names></name> <name><surname>Zakharko</surname> <given-names>T.</given-names></name> <name><surname>Witzlack-Makarevich</surname> <given-names>A.</given-names></name> <name><surname>Hildebrandt</surname> <given-names>K.</given-names></name> <name><surname>Rie&#x00DF;ler</surname> <given-names>M.</given-names></name> <etal/></person-group>. (<year>2017</year>). The AUTOTYP typological databases. Version 0.1.0. Available at: <ext-link xlink:href="https://zenodo.org/record/3667562#.YineCJYo9EY" ext-link-type="uri">https://zenodo.org/record/3667562#.YineCJYo9EY</ext-link></citation></ref>
<ref id="ref5"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Brown</surname> <given-names>C. H.</given-names></name> <name><surname>Holman</surname> <given-names>E. W.</given-names></name> <name><surname>Wichmann</surname> <given-names>S.</given-names></name></person-group> (<year>2013</year>). <article-title>Sound correspondences in the world&#x2019;s languages</article-title>. <source>Language</source> <volume>89</volume>, <fpage>4</fpage>&#x2013;<lpage>29</lpage>. doi: <pub-id pub-id-type="doi">10.1353/lan.2013.0009</pub-id>, PMID: <pub-id pub-id-type="pmid">35217676</pub-id></citation></ref>
<ref id="ref6"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dediu</surname> <given-names>D.</given-names></name></person-group> (<year>2018</year>). <article-title>Making genealogical language classifications available for phylogenetic analysis: Newick trees, unified identifiers, and branch length</article-title>. <source>Lang. Dyn. Chang.</source> <volume>8</volume>, <fpage>1</fpage>&#x2013;<lpage>21</lpage>. doi: <pub-id pub-id-type="doi">10.1163/22105832-00801001</pub-id></citation></ref>
<ref id="ref7"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dellert</surname> <given-names>J.</given-names></name> <name><surname>Daneyko</surname> <given-names>T.</given-names></name> <name><surname>M&#x00FC;nch</surname> <given-names>A.</given-names></name> <name><surname>Ladygina</surname> <given-names>A.</given-names></name> <name><surname>Buch</surname> <given-names>A.</given-names></name> <name><surname>Clarius</surname> <given-names>N.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>NorthEuraLex: a wide-coverage lexical database of Northern Eurasia</article-title>. <source>Lang. Resources Eval.</source> <volume>54</volume>, <fpage>273</fpage>&#x2013;<lpage>301</lpage>. doi: <pub-id pub-id-type="doi">10.1007/s10579-019-09480-6</pub-id>, PMID: <pub-id pub-id-type="pmid">32214931</pub-id></citation></ref>
<ref id="ref8"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Dray</surname> <given-names>S.</given-names></name> <name><surname>Dufour</surname> <given-names>A.</given-names></name></person-group> (<year>2007</year>). <article-title>The ade4 package: implementing the duality diagram for ecologists</article-title>. <source>J. Stat. Softw.</source> <volume>22</volume>, <fpage>1</fpage>&#x2013;<lpage>20</lpage>. doi: <pub-id pub-id-type="doi">10.18637/jss.v022.i04</pub-id></citation></ref>
<ref id="ref9"><citation citation-type="book"><person-group person-group-type="editor"><name><surname>Dryer</surname> <given-names>M. S.</given-names></name> <name><surname>Haspelmath</surname> <given-names>M</given-names></name></person-group>. Eds. (<year>2013</year>). <source>The world atlas of language structures online</source>. (<publisher-loc>Leipzig</publisher-loc>: <publisher-name>Max Planck Institute for Evolutionary Anthropology</publisher-name>).</citation></ref>
<ref id="ref10"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Ferlus</surname> <given-names>M.</given-names></name></person-group> (<year>2009</year>). <article-title>What were the four divisions of middle Chinese?</article-title> <source>Diachronica</source> <volume>26</volume>, <fpage>184</fpage>&#x2013;<lpage>213</lpage>. doi: <pub-id pub-id-type="doi">10.1075/dia.26.2.02fer</pub-id>, PMID: <pub-id pub-id-type="pmid">36971353</pub-id></citation></ref>
<ref id="ref11"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Hammarstr&#x00F6;m</surname> <given-names>H.</given-names></name> <name><surname>Forkel</surname> <given-names>R.</given-names></name> <name><surname>Haspelmath</surname> <given-names>M.</given-names></name> <name><surname>Bank</surname> <given-names>S</given-names></name></person-group>. (<year>2021</year>). <source>Glottolog 4.4</source>. <publisher-loc>Leipzig</publisher-loc>: <publisher-name>Max Planck Institute for Evolutionary Anthropology</publisher-name>.</citation></ref>
<ref id="ref12"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Harvey</surname> <given-names>P. H.</given-names></name> <name><surname>Pagel</surname> <given-names>M. D.</given-names></name></person-group> (<year>1991</year>). <source>The comparative method in evolutionary biology</source>. (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>).</citation></ref>
<ref id="ref13"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Haudricourt</surname> <given-names>A.-H.</given-names></name></person-group> (<year>1954</year>). <article-title>De l&#x2019;origine des tons en vietnamien</article-title>. <source>J. Asiat.</source> <volume>242</volume>, <fpage>69</fpage>&#x2013;<lpage>82</lpage>.</citation></ref>
<ref id="ref14"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Holman</surname> <given-names>E. W.</given-names></name> <name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>Brown</surname> <given-names>C. H.</given-names></name> <name><surname>Velupillai</surname> <given-names>V.</given-names></name> <name><surname>M&#x00FC;ller</surname> <given-names>A.</given-names></name> <name><surname>Bakker</surname> <given-names>D.</given-names></name></person-group> (<year>2008</year>). &#x201C;<article-title>Advances in automated language classification</article-title>&#x201D; in <source>Quantitative investigations in theoretical linguistics</source>. eds. <person-group person-group-type="editor"><name><surname>Arppe</surname> <given-names>A.</given-names></name> <name><surname>Sinnem&#x00E4;ki</surname> <given-names>K.</given-names></name> <name><surname>Nikanne</surname> <given-names>U.</given-names></name></person-group> (<publisher-loc>Helsinki</publisher-loc>: <publisher-name>University of Helsinki</publisher-name>), <fpage>40</fpage>&#x2013;<lpage>43</lpage>.</citation></ref>
<ref id="ref15"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>J&#x00E4;ger</surname> <given-names>G.</given-names></name></person-group> (<year>2018</year>). <article-title>Global-scale phylogenetics linguistic inference from lexical resources</article-title>. <source>Sci. Data</source> <volume>5</volume>:<fpage>180189</fpage>. doi: <pub-id pub-id-type="doi">10.1038/sdata.2018.189</pub-id>, PMID: <pub-id pub-id-type="pmid">30299438</pub-id></citation></ref>
<ref id="ref16"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Maddieson</surname> <given-names>I.</given-names></name></person-group> (<year>2007</year>). &#x201C;<article-title>Issues of phonological complexity: statistical analysis of the relationship between syllable structures, segment inventories, and tone contrasts</article-title>&#x201D; in <source>Experimental approaches to phonology</source>. eds. <person-group person-group-type="editor"><name><surname>Sol&#x00E9;</surname> <given-names>M.-J.</given-names></name> <name><surname>Beddor</surname> <given-names>P. S.</given-names></name> <name><surname>Ohala</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>93</fpage>&#x2013;<lpage>103</lpage>.</citation></ref>
<ref id="ref17"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Maddieson</surname> <given-names>I.</given-names></name></person-group> (<year>2013</year>). &#x201C;<article-title>Tone</article-title>&#x201D; in <source>The world atlas of language structures online</source>. eds. <person-group person-group-type="editor"><name><surname>Dryer</surname> <given-names>M. S.</given-names></name> <name><surname>Haspelmath</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Leipzig</publisher-loc>: <publisher-name>Max Planck Institute for Evolutionary Anthropology</publisher-name>)</citation></ref>
<ref id="ref18"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Meade</surname> <given-names>A.</given-names></name> <name><surname>Pagel</surname> <given-names>M.</given-names></name></person-group> (<year>2023</year>). BayesTraits V4.0.1. Software and manual. Available at: <ext-link xlink:href="http://www.evolution.reading.ac.uk/BayesTraitsV4.0.1/BayesTraitsV4.0.1.html" ext-link-type="uri">http://www.evolution.reading.ac.uk/BayesTraitsV4.0.1/BayesTraitsV4.0.1.html</ext-link>.</citation></ref>
<ref id="ref19"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Michaud</surname> <given-names>A.</given-names></name> <name><surname>Sands</surname> <given-names>B.</given-names></name></person-group> (<year>2020</year>). &#x201C;<article-title>Tonogenesis</article-title>&#x201D; in <source>Oxford research Encyclopedia of linguistics</source>. ed. <person-group person-group-type="editor"><name><surname>Aronoff</surname> <given-names>M.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>)</citation></ref>
<ref id="ref20"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Moran</surname> <given-names>S. P.</given-names></name></person-group> (<year>2012</year>). <source>Phonetics information base and lexicon</source>. <comment>Ph.D. Dissertation</comment>. <publisher-loc>Seattle, WA</publisher-loc>: <publisher-name>University of Washington</publisher-name>.</citation></ref>
<ref id="ref21"><citation citation-type="confproc"><person-group person-group-type="author"><name><surname>Moran</surname> <given-names>S.</given-names></name> <name><surname>Bentz</surname> <given-names>C.</given-names></name> <name><surname>Gutierrez-Vasques</surname> <given-names>X.</given-names></name> <name><surname>Sozinova</surname> <given-names>O.</given-names></name> <name><surname>Samardzic</surname> <given-names>T.</given-names></name></person-group> (<year>2022</year>). <article-title>TeDDi sample: text data diversity sample for language comparison and multilingual NLP</article-title>. <conf-name>Proceedings of the 13th Conference on Language Resources and Evaluation (LREC 2022)</conf-name>, <conf-loc>Marseille</conf-loc>, <conf-date>20&#x2013;25 June 2022</conf-date>, <fpage>1150</fpage>&#x2013;<lpage>1158</lpage>.</citation></ref>
<ref id="ref22"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Moran</surname> <given-names>S.</given-names></name> <name><surname>Blasi</surname> <given-names>D.</given-names></name></person-group> (<year>2014</year>). &#x201C;<article-title>Cross-linguistic comparison of complexity measures in phonological systems</article-title>&#x201D; in <source>Measuring grammatical complexity</source>. eds. <person-group person-group-type="editor"><name><surname>Newmeyer</surname> <given-names>F. J.</given-names></name> <name><surname>Preston</surname> <given-names>L. B.</given-names></name></person-group> (<publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>), <fpage>217</fpage>&#x2013;<lpage>240</lpage>.</citation></ref>
<ref id="ref23"><citation citation-type="book"><person-group person-group-type="editor"><name><surname>Moran</surname> <given-names>S.</given-names></name> <name><surname>McCloy</surname> <given-names>D.</given-names></name></person-group> (Eds). (<year>2019</year>). <source>PHOIBLE 2.0</source>. <publisher-loc>Jena</publisher-loc>: <publisher-name>Max Planck Institute for the Science of Human History</publisher-name>.</citation></ref>
<ref id="ref24"><citation citation-type="other"><person-group person-group-type="author"><name><surname>M&#x00FC;ller</surname> <given-names>K.</given-names></name> <name><surname>Wickham</surname> <given-names>H.</given-names></name></person-group> (<year>2022</year>). tibble: Simple Data Frames. R package version 3.1.8. Available at: <ext-link xlink:href="https://CRAN.R-project.org/package=tibble" ext-link-type="uri">https://CRAN.R-project.org/package=tibble</ext-link>.</citation></ref>
<ref id="ref25"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nettle</surname> <given-names>D.</given-names></name></person-group> (<year>1995</year>). <article-title>Segmental inventory size, word length, and communicative efficiency</article-title>. <source>Linguistics</source> <volume>33</volume>, <fpage>359</fpage>&#x2013;<lpage>367</lpage>.</citation></ref>
<ref id="ref26"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nettle</surname> <given-names>D.</given-names></name></person-group> (<year>1998</year>). <article-title>Coevolution of phonology and the lexicon in twelve languages of West Africa</article-title>. <source>J. Quant. Linguist.</source> <volume>5</volume>, <fpage>240</fpage>&#x2013;<lpage>245</lpage>. doi: <pub-id pub-id-type="doi">10.1080/09296179808590132</pub-id></citation></ref>
<ref id="ref27"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Nikolaev</surname> <given-names>D.</given-names></name></person-group> (<year>2018</year>). <article-title>The database of Eurasian phonological inventories: a research tool for distributional phonological typology</article-title>. <source>Linguist. Vanguard</source> <volume>4</volume>:<fpage>20170050</fpage>. doi: <pub-id pub-id-type="doi">10.1515/lingvan-2017-0050</pub-id></citation></ref>
<ref id="ref28"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pagel</surname> <given-names>M.</given-names></name></person-group> (<year>1994</year>). <article-title>Detecting correlated evolution on phylogenies: a general method for the comparative analysis of discrete characters</article-title>. <source>Proc. R. Soc. Lond. B</source> <volume>255</volume>, <fpage>37</fpage>&#x2013;<lpage>45</lpage>. doi: <pub-id pub-id-type="doi">10.1098/rspb.1994.0006</pub-id></citation></ref>
<ref id="ref29"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pagel</surname> <given-names>M.</given-names></name></person-group> (<year>1997</year>). <article-title>Inferring evolutionary processes from phylogenies</article-title>. <source>Zool. Scripta</source> <volume>26</volume>, <fpage>331</fpage>&#x2013;<lpage>348</lpage>. doi: <pub-id pub-id-type="doi">10.1111/j.1463-6409.1997.tb00423.x</pub-id>, PMID: <pub-id pub-id-type="pmid">37097343</pub-id></citation></ref>
<ref id="ref30"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pagel</surname> <given-names>M.</given-names></name></person-group> (<year>1999</year>). <article-title>Inferring the historical patterns of biological evolution</article-title>. <source>Nature</source> <volume>401</volume>, <fpage>877</fpage>&#x2013;<lpage>884</lpage>. doi: <pub-id pub-id-type="doi">10.1038/44766</pub-id>, PMID: <pub-id pub-id-type="pmid">10553904</pub-id></citation></ref>
<ref id="ref31"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pagel</surname> <given-names>M.</given-names></name> <name><surname>Meade</surname> <given-names>A.</given-names></name> <name><surname>Barker</surname> <given-names>D.</given-names></name></person-group> (<year>2004</year>). <article-title>Bayesian estimation of ancestral character states on phylogenies</article-title>. <source>Syst. Biol.</source> <volume>53</volume>, <fpage>673</fpage>&#x2013;<lpage>684</lpage>. doi: <pub-id pub-id-type="doi">10.1080/10635150490522232</pub-id>, PMID: <pub-id pub-id-type="pmid">15545248</pub-id></citation></ref>
<ref id="ref32"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Paradis</surname> <given-names>E.</given-names></name> <name><surname>Schliep</surname> <given-names>K.</given-names></name></person-group> (<year>2019</year>). <article-title>Ape 5.0: an environment for modern phylogenetics and evolutionary analyses in R</article-title>. <source>Bioinformatics</source> <volume>35</volume>, <fpage>526</fpage>&#x2013;<lpage>528</lpage>. doi: <pub-id pub-id-type="doi">10.1093/bioinformatics/bty633</pub-id>, PMID: <pub-id pub-id-type="pmid">30016406</pub-id></citation></ref>
<ref id="ref33"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Pittayaporn</surname> <given-names>P.</given-names></name> <name><surname>Kirby</surname> <given-names>J.</given-names></name></person-group> (<year>2017</year>). <article-title>Laryngeal contrasts in the Tai dialect of Cao B&#x1EB1;ng</article-title>. <source>J. Int. Phon. Assoc.</source> <volume>47</volume>, <fpage>65</fpage>&#x2013;<lpage>85</lpage>. doi: <pub-id pub-id-type="doi">10.1017/S0025100316000293</pub-id></citation></ref>
<ref id="ref34"><citation citation-type="book"><person-group person-group-type="author"><collab id="coll1">R Core Team</collab></person-group> (<year>2022</year>). <source>R: A language and environment for statistical computing</source>. (<publisher-loc>Vienna, Austria</publisher-loc>: <publisher-name>R Foundation for Statistical Computing</publisher-name>).</citation></ref>
<ref id="ref35"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Raftery</surname> <given-names>A. E.</given-names></name></person-group> (<year>1996</year>). &#x201C;<article-title>Hypothesis testing and model selection</article-title>&#x201D; in <source>Markov chain Monte Carlo in practice</source>. eds. <person-group person-group-type="editor"><name><surname>Gilks</surname> <given-names>W. R.</given-names></name> <name><surname>Richardson</surname> <given-names>S.</given-names></name> <name><surname>Spiegelhalter</surname> <given-names>D. J.</given-names></name></person-group> (<publisher-loc>Dordrecht</publisher-loc>: <publisher-name>Springer Science + Business Media</publisher-name>), <fpage>163</fpage>&#x2013;<lpage>187</lpage>.</citation></ref>
<ref id="ref36"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Revell</surname> <given-names>L.</given-names></name></person-group> (<year>2011</year>). For fun: least squares phylogeny estimation. Blog post. Available at: <ext-link xlink:href="http://blog.phytools.org/2011/03/for-fun-least-squares-phylogeny.html" ext-link-type="uri">http://blog.phytools.org/2011/03/for-fun-least-squares-phylogeny.html</ext-link></citation></ref>
<ref id="ref37"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Round</surname> <given-names>E. R.</given-names></name></person-group> (<year>2021</year>). glottoTrees: phylogenetic trees in linguistics. R package version 0.1. Available at: <ext-link xlink:href="https://github.com/erichround/glottoTrees" ext-link-type="uri">https://github.com/erichround/glottoTrees</ext-link>.</citation></ref>
<ref id="ref38"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Sagart</surname> <given-names>L.</given-names></name></person-group> (<year>1999</year>). &#x201C;<article-title>The origin of Chinese tones</article-title>&#x201D; in <source>Cross-linguistic studies of tonal phenomena: tonogenesis, typology, and related topics</source>. ed. <person-group person-group-type="editor"><name><surname>Kaji</surname> <given-names>S.</given-names></name></person-group> (<publisher-loc>Tokyo</publisher-loc>: <publisher-name>ILCAA</publisher-name>), <fpage>91</fpage>&#x2013;<lpage>103</lpage>.</citation></ref>
<ref id="ref39"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Schliep</surname> <given-names>K. P.</given-names></name></person-group> (<year>2011</year>). <article-title>phangorn: phylogenetic analysis in R</article-title>. <source>Bioinformatics</source> <volume>27</volume>, <fpage>592</fpage>&#x2013;<lpage>593</lpage>. doi: <pub-id pub-id-type="doi">10.1093/bioinformatics/btq706</pub-id>, PMID: <pub-id pub-id-type="pmid">21169378</pub-id></citation></ref>
<ref id="ref40"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Shcherbakova</surname> <given-names>O.</given-names></name> <name><surname>Gast</surname> <given-names>V.</given-names></name> <name><surname>Blasi</surname> <given-names>D. E.</given-names></name> <name><surname>Skirg&#x00E5;rd</surname> <given-names>H.</given-names></name> <name><surname>Gray</surname> <given-names>R. D.</given-names></name> <name><surname>Greenhill</surname> <given-names>S. J.</given-names></name></person-group> (<year>2022</year>). <article-title>A quantitative global test of the complexity trade-off hypothesis: the case of nominal and verbal grammatical marking</article-title>. <source>Linguistics Vanguard</source>. doi: <pub-id pub-id-type="doi">10.1515/lingvan-2021-0011</pub-id></citation></ref>
<ref id="ref41"><citation citation-type="book"><person-group person-group-type="editor"><name><surname>Simons</surname> <given-names>G. F.</given-names></name> <name><surname>Fennig</surname> <given-names>C. D.</given-names></name></person-group> (Eds.) (<year>2017</year>). <source>Ethnologue: languages of the world</source>, <edition>20th</edition>. (<publisher-loc>Dallas, TX</publisher-loc>: <publisher-name>SIL International</publisher-name>).</citation></ref>
<ref id="ref42"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>South</surname> <given-names>A.</given-names></name></person-group> (<year>2011</year>). <article-title>rworldmap: a new R package for mapping global data</article-title>. <source>R. J.</source> <volume>3</volume>, <fpage>35</fpage>&#x2013;<lpage>43</lpage>. doi: <pub-id pub-id-type="doi">10.32614/RJ-2011-006</pub-id></citation></ref>
<ref id="ref43"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Tamura</surname> <given-names>K.</given-names></name> <name><surname>Stecher</surname> <given-names>G.</given-names></name> <name><surname>Kumar</surname> <given-names>S.</given-names></name></person-group> (<year>2021</year>). <article-title>MEGA11: molecular evolutionary genetics analysis version 11</article-title>. <source>Mol. Biol. Evol.</source> <volume>38</volume>, <fpage>3022</fpage>&#x2013;<lpage>3027</lpage>. doi: <pub-id pub-id-type="doi">10.1093/molbev/msab120</pub-id>, PMID: <pub-id pub-id-type="pmid">33892491</pub-id></citation></ref>
<ref id="ref44"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name></person-group> (<year>2013</year>). &#x201C;<article-title>A classification of Papuan languages</article-title>&#x201D; in <source>History, contact and classification of Papuan languages language and linguistics in Melanesia</source>, <comment>Special Issue 2012</comment>. eds. <person-group person-group-type="editor"><name><surname>Hammarstr&#x00F6;m</surname> <given-names>H.</given-names></name> <name><surname>van den Heuvel</surname> <given-names>W.</given-names></name></person-group> (<publisher-loc>Port Moresby</publisher-loc>: <publisher-name>Linguistic Society of Papua New Guinea</publisher-name>), <fpage>313</fpage>&#x2013;<lpage>386</lpage>.</citation></ref>
<ref id="ref45"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name></person-group> (<year>2023</year>). InteractiveASJP02. Software available at: <ext-link xlink:href="https://github.com/Sokiwi/InteractiveASJP02" ext-link-type="uri">https://github.com/Sokiwi/InteractiveASJP02</ext-link>.</citation></ref>
<ref id="ref46"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>Holman</surname> <given-names>E. W.</given-names></name></person-group> (<year>2023</year>). <article-title>Cross-linguistic conditions on word length</article-title>. <source>PLoS One</source> <volume>18</volume>:<fpage>e0281041</fpage>. doi: <pub-id pub-id-type="doi">10.1371/journal.pone.0281041</pub-id>, PMID: <pub-id pub-id-type="pmid">36706125</pub-id></citation></ref>
<ref id="ref47"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>Holman</surname> <given-names>E. W.</given-names></name> <name><surname>Bakker</surname> <given-names>D.</given-names></name> <name><surname>Brown</surname> <given-names>C. H.</given-names></name></person-group> (<year>2010a</year>). <article-title>Evaluating linguistic distance measures</article-title>. <source>Physica A</source> <volume>389</volume>, <fpage>3632</fpage>&#x2013;<lpage>3639</lpage>. doi: <pub-id pub-id-type="doi">10.1016/j.physa.2010.05.011</pub-id></citation></ref>
<ref id="ref70"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>Holman</surname> <given-names>E. W.</given-names></name> <name><surname>Brown</surname> <given-names>C. H.</given-names></name></person-group> (Eds.) (<year>2022</year>) The ASJP database (version 20). Available at: <ext-link xlink:href="https://asjp.clld.org/" ext-link-type="uri">https://asjp.clld.org/</ext-link>.</citation></ref>
<ref id="ref48"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>Holman</surname> <given-names>E. W.</given-names></name> <name><surname>Brown</surname> <given-names>C. H.</given-names></name></person-group> (Eds.) (<year>2020</year>) The ASJP database (version 19). Available at: <ext-link xlink:href="https://asjp.clld.org/" ext-link-type="uri">https://asjp.clld.org/</ext-link>.</citation></ref>
<ref id="ref49"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>M&#x00FC;ller</surname> <given-names>A.</given-names></name> <name><surname>Velupillai</surname> <given-names>V.</given-names></name></person-group> (<year>2010b</year>). <article-title>Homelands of the world&#x2019;s language families: a quantitative approach</article-title>. <source>Diachronica</source> <volume>27</volume>, <fpage>247</fpage>&#x2013;<lpage>276</lpage>. doi: <pub-id pub-id-type="doi">10.1075/dia.27.2.05wic</pub-id></citation></ref>
<ref id="ref50"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Wichmann</surname> <given-names>S.</given-names></name> <name><surname>Rama</surname> <given-names>T.</given-names></name> <name><surname>Holman</surname> <given-names>E. W.</given-names></name></person-group> (<year>2011</year>). <article-title>Phonological diversity, word length, and population sizes across languages: the ASJP evidence</article-title>. <source>Linguist. Typol.</source> <volume>15</volume>, <fpage>177</fpage>&#x2013;<lpage>197</lpage>. doi: <pub-id pub-id-type="doi">10.1515/LITY.2011.013</pub-id></citation></ref>
<ref id="ref51"><citation citation-type="book"><person-group person-group-type="author"><name><surname>Wickham</surname> <given-names>H.</given-names></name></person-group> (<year>2016</year>). <source>ggplot2: elegant graphics for data analysis</source>. (<publisher-loc>New York</publisher-loc>: <publisher-name>Springer-Verlag</publisher-name>).</citation></ref>
<ref id="ref52"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Wickham</surname> <given-names>H.</given-names></name></person-group> (<year>2022</year>). stringr: simple, consistent wrappers for common string operations. R package version 1.5.0. Available at: <ext-link xlink:href="https://CRAN.R-project.org/package=stringr%3e" ext-link-type="uri">https://CRAN.R-project.org/package=stringr</ext-link>.</citation></ref>
<ref id="ref53"><citation citation-type="other"><person-group person-group-type="author"><name><surname>Wickham</surname> <given-names>H</given-names></name> <name><surname>Fran&#x00E7;ois</surname> <given-names>R</given-names></name></person-group>., <person-group person-group-type="author"><name><surname>Henry</surname> <given-names>L</given-names></name></person-group>., <person-group person-group-type="author"><name><surname>M&#x00FC;ller</surname> <given-names>K</given-names></name></person-group>., and <person-group person-group-type="author"><name><surname>Vaughan</surname> <given-names>D</given-names></name></person-group>. (<year>2023</year>). dplyr: a grammar of data manipulation. R package version 1.1.0. Available at: <ext-link xlink:href="https://CRAN.R-project.org/package=dplyr" ext-link-type="uri">https://CRAN.R-project.org/package=dplyr</ext-link>.</citation></ref>
<ref id="ref54"><citation citation-type="journal"><person-group person-group-type="author"><name><surname>Zeileis</surname> <given-names>A.</given-names></name> <name><surname>Fisher</surname> <given-names>J. C.</given-names></name> <name><surname>Hornik</surname> <given-names>K.</given-names></name> <name><surname>Ihaka</surname> <given-names>R.</given-names></name> <name><surname>McWhite</surname> <given-names>C. D.</given-names></name> <name><surname>Murrell</surname> <given-names>P.</given-names></name> <etal/></person-group>. (<year>2020</year>). <article-title>colorspace: a toolbox for manipulating and assessing colors and palettes</article-title>. <source>J. Stat. Softw.</source> <volume>96</volume>, <fpage>1</fpage>&#x2013;<lpage>49</lpage>. doi: <pub-id pub-id-type="doi">10.18637/jss.v096.i01</pub-id></citation></ref>
</ref-list>
<fn-group>
<fn id="fn0003"><p><sup>1</sup><ext-link xlink:href="https://wals.info/languoid/samples/100" ext-link-type="uri">https://wals.info/languoid/samples/100</ext-link></p></fn><fn id="fn0004"><p><sup>2</sup><ext-link xlink:href="https://phoible.org/contributors/PH" ext-link-type="uri">https://phoible.org/contributors/PH</ext-link></p></fn><fn id="fn0005"><p><sup>3</sup><ext-link xlink:href="https://evolution.genetics.washington.edu/phylip/newicktree.html" ext-link-type="uri">https://evolution.genetics.washington.edu/phylip/newicktree.html</ext-link>, for instance.</p></fn>
</fn-group>
</back>
</article>