<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.0 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.0/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.0" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology: Journal of the Association for Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/labphon.100</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Statistical and acoustic effects on the perception of stop consonants in Kaqchikel (Mayan)</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<contrib-id contrib-id-type="orcid">http://orcid.org/0000-0001-6160-7007</contrib-id>
<name>
<surname>Bennett</surname>
<given-names>Ryan</given-names>
</name>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<contrib-id contrib-id-type="orcid">http://orcid.org/0000-0001-7382-9344</contrib-id>
<name>
<surname>Tang</surname>
<given-names>Kevin</given-names>
</name>
<email>linguist@kevintang.org</email>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Sian</surname>
<given-names>Juan Ajsivinac</given-names>
</name>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>Department of Linguistics, University of California, Santa Cruz, CA, US</aff>
<aff id="aff-2"><label>2</label>Department of Linguistics, Zhejiang University, Hangzhou, 310058, CN</aff>
<aff id="aff-3"><label>3</label>Independent, GT</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2018-05-25">
<day>25</day>
<month>05</month>
<year>2018</year>
</pub-date>
<pub-date pub-type="collection">
<year>2018</year>
</pub-date>
<volume>9</volume>
<issue>1</issue>
<elocation-id>9</elocation-id>
<history>
<date date-type="received" iso-8601-date="2017-06-26">
<day>26</day>
<month>06</month>
<year>2017</year>
</date>
<date date-type="accepted" iso-8601-date="2018-03-19">
<day>19</day>
<month>03</month>
<year>2018</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2018 The Author(s)</copyright-statement>
<copyright-year>2018</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.5334/labphon.100/"/>
<abstract>
<p>This paper investigates the relationship between speech perception and linguistic experience in Kaqchikel, a Guatemalan Mayan language. Our empirical focus is the perception of plain, ejective, and implosive stops. Drawing on an AX discrimination task, a corpus of spoken Kaqchikel, and a text corpus, we make two claims. First, we argue that speech perception is mediated by phonemic representations which include acoustic detail drawn from prior phonetic experience, as in Exemplar Theory. Second, segmental distributions also condition speech perception: The perceptual distinctiveness of a pair of phonemes is affected by their functional load and relative contextual predictability. These top-down factors influence phoneme discrimination even at relatively fast response times. We take this result as evidence that distributional factors like functional load may affect speech perception by shaping perceptual tuning during linguistic development. This study replicates and extends some key findings in speech perception in the context of a language (Kaqchikel) which is structurally and sociolinguistically different from the majority languages (like English) which have served as the basis of most work in the speech perception literature. At the practical level, our research illustrates methods for conducting corpus-based laboratory phonology with lesser-studied and under-resourced languages.</p>
</abstract>
<kwd-group>
<kwd>contrast</kwd>
<kwd>discriminability</kwd>
<kwd>functional load</kwd>
<kwd>Exemplar Theory</kwd>
<kwd>Mayan languages</kwd>
<kwd>ejectives</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>1. Speech perception and linguistic experience</title>
<p>The perceptual similarity of any two speech sounds depends, to a great extent, on their raw acoustic similarity. However, speech perception is also mediated by the native language(s) of the hearer. The functional organization of speech sounds within a language&#8217;s phonological system strongly determines whether a given pair of segments will be well-discriminated by speakers of that language (<xref ref-type="bibr" rid="B176">Trubetzkoy, 1939</xref>; see <xref ref-type="bibr" rid="B14">Best, 1995</xref>; <xref ref-type="bibr" rid="B15">Best, McRoberts, &amp; Goodell, 2001</xref>; <xref ref-type="bibr" rid="B19">Boomershine, Hall, Hume, &amp; Johnson, 2008</xref>; <xref ref-type="bibr" rid="B159">Sebasti&#225;n-Gall&#233;s, 2005</xref> for references and recent discussion). For example, [d &#240;] are present in both English and in Spanish. In English these sounds are contrastive, as attested by minimal pairs like /be&#865;&#618;d/ &#8216;bade&#8217; vs. /be&#865;&#618;&#240;/ &#8216;bathe.&#8217; In Spanish [&#240;] is instead a conditioned allophone of the phoneme /d/, e.g., [de&#240;o] &#8216;finger&#8217; vs. [ese &#240;e&#240;o] &#8216;that finger&#8217; (e.g., <xref ref-type="bibr" rid="B87">Harris, 1969</xref>). This difference in the function of [d &#240;] has consequences for speech perception: English speakers, who rely on the [d &#240;] contrast to distinguish word meanings, are better at discriminating these sounds than Spanish speakers, for whom [&#240;] is simply a predictable variant of /d/ (<xref ref-type="bibr" rid="B19">Boomershine et al., 2008</xref>; see too <xref ref-type="bibr" rid="B84">Harnsberger, 2000</xref>, <xref ref-type="bibr" rid="B85">2001a</xref>, <xref ref-type="bibr" rid="B86">2001b</xref>). These and related findings demonstrate that prior linguistic experience plays a significant role in conditioning the perception of speech sounds.</p>
<p>Even fine-grained details of linguistic experience, based on statistical properties of a hearer&#8217;s native language, may substantially impact speech perception (see <xref ref-type="bibr" rid="B38">Cutler, 2012</xref> for an overview). Phoneme discrimination, for instance, appears to be sensitive to the specific acoustic parameters associated with each phoneme category in the hearer&#8217;s language (e.g., <xref ref-type="bibr" rid="B109">Kuhl &amp; Iverson, 1995</xref> and Section 5.2). More recent research has suggested that the statistical structure of the lexicon may also influence native language speech perception (e.g., <xref ref-type="bibr" rid="B81">Hall &amp; Hume, in preparation</xref>; see <xref ref-type="bibr" rid="B82">Hall, Hume, Jaeger, &amp; Wedel, in preparation</xref>; <xref ref-type="bibr" rid="B83">Hall, Letawsky, Turner, Allen, &amp; McMullin, 2014</xref>; <xref ref-type="bibr" rid="B99">Kataoka &amp; Johnson, 2007</xref>; <xref ref-type="bibr" rid="B179">Vitevitch &amp; Luce, 2016</xref> and references there; <xref ref-type="bibr" rid="B192">Yao, 2011</xref>, Ch.2 for related discussion). Among other factors, segment-level measures such as phoneme frequency and functional load (discussed in Section 5.3) seem to contribute to the relative discriminability of different phoneme pairs. The precise mechanism behind such effects is not well-understood at present, a point we return to in Section 7.</p>
<p>It thus seems clear that statistical properties of a hearer&#8217;s native language may influence speech perception. However, we believe that the full generality of these findings has not yet been established, particularly with respect to lexical effects on speech perception. A large proportion of speech perception studies&#8212;perhaps most such studies&#8212;involve experiments with listeners who are native speakers of majority languages like English, French, Dutch, Japanese, and so on (e.g., <xref ref-type="bibr" rid="B38">Cutler, 2012, p. 4</xref>). This sampling bias might be unremarkable, if not for the fact that the languages most commonly used in speech perception research also share a number of structural and sociolinguistic properties. To give one example, the European languages most often used in speech perception research typically belong to the Germanic or Romance branches of the Indo-European family. The morphological structure of these languages is characteristically analytic or fusional rather than agglutinating. Since perfect minimal pairs should, intuitively, be less common in languages which tend toward longer and/or more complex words, it remains unclear whether statistical measures which refer to minimal pairs (e.g., functional load) should have the same importance in languages with relatively agglutinative morphology (see also <xref ref-type="bibr" rid="B82">Hall et al., in preparation</xref>; <xref ref-type="bibr" rid="B182">Wedel, Jackson, &amp; Kaplan, 2013</xref>; <xref ref-type="bibr" rid="B183">Wedel, Kaplan, &amp; Jackson, 2013</xref>).<xref ref-type="fn" rid="n1">1</xref> Similar questions arise with respect to neighborhood density and word-frequency effects in agglutinating languages, as longer words are likely to have fewer lexical neighbors and (possibly) low overall corpus frequencies (e.g., <xref ref-type="bibr" rid="B179">Vitevitch &amp; Luce, 2016</xref>; <xref ref-type="bibr" rid="B192">Yao, 2011</xref>; <xref ref-type="bibr" rid="B195">Zipf, 1935</xref>, Ch.2).</p>
<p>At the sociolinguistic level, a preponderance of work in speech perception has been conducted with listeners who are highly educated and literate in their native language. Apart from general concerns about whether results obtained with such populations are really generalizable (e.g., <xref ref-type="bibr" rid="B90">Henrich, Heine, &amp; Norenzayan, 2010</xref>), the bias in speech perception studies toward literate speakers of Indo-European languages is potentially relevant for understanding how phoneme-level lexical statistics interact with speech perception. It has sometimes been suggested that lexical items lack a phoneme-level encoding altogether, being stored instead with strictly gestural and/or syllabic encoding (e.g., <xref ref-type="bibr" rid="B22">Browman &amp; Goldstein, 1986</xref>, <xref ref-type="bibr" rid="B23">1989</xref>, <xref ref-type="bibr" rid="B24">1992</xref>, etc.; <xref ref-type="bibr" rid="B112">Ladefoged &amp; Disner, 2012</xref>; <xref ref-type="bibr" rid="B115">Lodge, 2009</xref>; <xref ref-type="bibr" rid="B149">Port &amp; Leary, 2005</xref>; <xref ref-type="bibr" rid="B161">Silverman, 2006</xref>, <xref ref-type="bibr" rid="B162">2012</xref>; <xref ref-type="bibr" rid="B175">Tilsen, 2016</xref>; cf. <xref ref-type="bibr" rid="B50">Dunbar &amp; Idsardi, 2010</xref>; <xref ref-type="bibr" rid="B94">Hyman, 2015</xref>). To the extent that such models of lexical storage can account for phoneme-level statistical effects in speech perception, they would presumably attribute such effects to phonemic awareness, itself an artifact of literacy in an alphabetic writing system. Against this backdrop, further studies of speech perception among populations with non-alphabetic writing systems, or simply low literacy rates, are clearly needed. It is not our place here to adjudicate between these views, only to highlight the fact that answering such questions will require a more diverse sample of speakers and languages than currently exists in the speech perception literature.</p>
<p>In this article we explore how statistical measures derived from the lexicon (such as pairwise functional load) affect stop consonant discrimination in Kaqchikel, a Guatemalan Mayan language (Section 2). Kaqchikel has a number of properties, both grammatical and sociolinguistic, which differentiate it from most of the majority languages typically encountered in the speech perception literature. As discussed in Section 9, our study replicates some past results regarding the influence of segment-level distributional statistics on speech perception, and in doing so, supports the generality of such effects across different linguistic populations.</p>
<p>We investigate these issues using an AX discrimination study of the stop consonants of Kaqchikel (Section 3). Our emphasis in this paper is the influence of linguistic experience on speech perception in a lesser-studied language. For reasons of space we do not discuss specific patterns of pairwise consonant confusion in detail here.</p>
</sec>
<sec>
<title>2. Kaqchikel</title>
<p>Kaqchikel is a K&#8217;ichean-branch Mayan language spoken by over half a million people in southern Guatemala (Figure <xref ref-type="fig" rid="F1">1</xref>; <xref ref-type="bibr" rid="B59">Fischer &amp; Brown, 1996</xref>, fn. 3; <xref ref-type="bibr" rid="B124">Maxwell &amp; Hill, 2010</xref>; <xref ref-type="bibr" rid="B155">Richards, 2003</xref>). Like all Mayan languages, Kaqchikel has a phonemic contrast between plain voiceless plosives (/p t k q t&#865;s t&#865;&#643;/) and &#8216;glottalized&#8217; plosives at corresponding places of articulation (implosive /&#595;/, ejective /t<sup>&#660;</sup> k<sup>&#660;</sup> q<sup>&#660;</sup> t&#865;s<sup>&#660;</sup> t&#865;&#643;<sup>&#660;</sup>/ and /&#660;/) (Table <xref ref-type="table" rid="T1">1</xref>; <xref ref-type="bibr" rid="B10">Bennett, 2016</xref>; <xref ref-type="bibr" rid="B32">Campbell, 1977</xref>; <xref ref-type="bibr" rid="B33">Chacach Cutzal, 1990</xref>; <xref ref-type="bibr" rid="B35">Cojt&#237; Macario &amp; Lopez, 1990</xref>; <xref ref-type="bibr" rid="B70">Garc&#237;a Matzar, Toj Cotzajay, &amp; Coc Tuiz, 1999</xref>; <xref ref-type="bibr" rid="B122">Majzul, Matzar, &amp; Serech, 2000</xref>; <xref ref-type="bibr" rid="B26">R. M. Brown, Maxwell, &amp; Little, 2010</xref>, etc.).</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Map of Guatemala showing the four administrative departments in which Kaqchikel is most widely spoken as a community language (from east to west, these are the departments of Guatemala, Sacatep&#233;quez, Chimaltenango, and Solol&#225;) (<xref ref-type="bibr" rid="B26">R. M. Brown et al., 2010</xref>; <xref ref-type="bibr" rid="B124">Maxwell &amp; Hill, 2010</xref>; <xref ref-type="bibr" rid="B155">Richards, 2003</xref>).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6226/file/75170/"/>
</fig>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>The phonemic consonants of Kaqchikel.</p>
</caption>
<table>
<tr>
<th align="left"></th>
<th align="center">Bilabial</th>
<th align="center">Dental/alveolar</th>
<th align="center">Postalveolar</th>
<th align="center">Velar</th>
<th align="center">Uvular</th>
<th align="center">Glottal</th>
</tr>
<tr>
<td colspan="7"><hr/></td>
</tr>
<tr>
<td align="left"><bold>Stop</bold></td>
<td align="center">p &#595;</td>
<td align="center">t t<sup>&#660;</sup></td>
<td align="center"></td>
<td align="center">k k<sup>&#660;</sup></td>
<td align="center">q q<sup>&#660;</sup></td>
<td align="center">&#660;</td>
</tr>
<tr>
<td align="left"><bold>Affricate</bold></td>
<td align="center"></td>
<td align="center">t&#865;s t&#865;s<sup>&#660;</sup></td>
<td align="center">t&#865;&#643; t&#865;&#643;<sup>&#660;</sup></td>
<td align="center"></td>
<td align="center"></td>
<td align="center"></td>
</tr>
<tr>
<td align="left"><bold>Fricative</bold></td>
<td align="center"></td>
<td align="center">s</td>
<td align="center">&#643;</td>
<td align="center" colspan="2">x &#126; &#967;</td>
<td align="center"></td>
</tr>
<tr>
<td align="left"><bold>Nasal</bold></td>
<td align="center">m</td>
<td align="center">n</td>
<td align="center"></td>
<td align="center"></td>
<td align="center"></td>
<td align="center"></td>
</tr>
<tr>
<td align="left"><bold>Semivowel</bold></td>
<td align="center">w</td>
<td align="center"></td>
<td align="center">j</td>
<td align="center"></td>
<td align="center"></td>
<td align="center"></td>
</tr>
<tr>
<td align="left"><bold>Liquid</bold></td>
<td align="center"></td>
<td align="center">l r</td>
<td align="center"></td>
<td align="center"></td>
<td align="center"></td>
<td align="center"></td>
</tr>
</table>
</table-wrap>
<p>Ejectives and implosives are reasonably uncommon cross-linguistically: Maddieson&#8217;s (<xref ref-type="bibr" rid="B119">2009</xref>) typological survey of 566 languages finds that 151 (27%) have either ejectives or implosives in their consonant inventories. Bennett, Tang, and Ajsivinac (<xref ref-type="bibr" rid="B13">in preparation</xref>) find that the ejectives of Kaqchikel closely resemble the &#8216;slack&#8217; ejectives described by Lindau (<xref ref-type="bibr" rid="B114">1984</xref>) and Kingston (<xref ref-type="bibr" rid="B105">1984</xref>, <xref ref-type="bibr" rid="B107">2005b</xref>): They are characteristically produced with short VOTs and weak release bursts, and cause creaky voice on adjacent voiced segments. Ejectives are sometimes realized with glottal closure following the oral release burst: This difference in release quality, along with creakiness in adjacent segments, seems to be a reliable cue to the plain&#126;ejective distinction in Kaqchikel. Implosive /&#595;/ usually lacks a release burst entirely, being ingressive, and also conditions creaky voice on neighboring sounds (<xref ref-type="bibr" rid="B10">Bennett, 2016</xref>; <xref ref-type="bibr" rid="B122">Majzul et al., 2000</xref>). Common realizations of /&#595;/ include [&#595;&#816; &#595;&#805; &#595;] less common realizations include [p<sup>&#660;</sup> w&#816; &#660;]. The phonetic realization of /q<sup>&#660;</sup>/ is typically either [q<sup>&#660;</sup>] or [&#667;&#805;]. These findings are all in line with past descriptions of Kaqchikel and other Eastern Mayan languages (e.g., <xref ref-type="bibr" rid="B5">Barrett, 1999</xref>; <xref ref-type="bibr" rid="B9">Bennett, 2010</xref>; <xref ref-type="bibr" rid="B49">DuBois, 1981</xref>; <xref ref-type="bibr" rid="B52">England, 1983</xref>; <xref ref-type="bibr" rid="B105">Kingston, 1984</xref>; <xref ref-type="bibr" rid="B113">Larsen, 1988</xref>; <xref ref-type="bibr" rid="B144">Pinkerton, 1986</xref>; <xref ref-type="bibr" rid="B157">Russell, 1997</xref>). Since not much previous research has been done on the perception of glottalized consonants in any language, and none at all on Mayan languages, we do not have any prior expectations as to how discriminable plain&#126;glottalized contrasts might be in Kaqchikel (on the perception of glottalized consonants outside of Mayan, see <xref ref-type="bibr" rid="B61">Fre Woldu, 1985</xref>; <xref ref-type="bibr" rid="B64">Gallagher, 2010a</xref>, <xref ref-type="bibr" rid="B65">2010b</xref>, <xref ref-type="bibr" rid="B66">2011</xref>, <xref ref-type="bibr" rid="B67">2012</xref>, <xref ref-type="bibr" rid="B68">2014</xref>; <xref ref-type="bibr" rid="B156">Rose &amp; King, 2007</xref>; <xref ref-type="bibr" rid="B190">R. Wright, Hargus, &amp; Davis, 2002</xref>).</p>
<p>The morphological system of Kaqchikel is moderately agglutinating, especially with verbs (see <xref ref-type="bibr" rid="B26">R. M. Brown et al., 2010</xref>; <xref ref-type="bibr" rid="B33">Chacach Cutzal, 1990</xref>; <xref ref-type="bibr" rid="B36">Coon, 2016</xref>; <xref ref-type="bibr" rid="B70">Garc&#237;a Matzar et al., 1999</xref>; <xref ref-type="bibr" rid="B100">Kaufman, 1990</xref>). Across lexical categories, the prefixal field is mostly reserved for inflectional affixes marking aspect and person/number agreement, while the suffixal field is composed of derivational affixes (1, 2) (the adjectival root <italic>ch&#8217;u&#8217;j</italic> /t&#865;&#643;<sup>&#660;</sup>u&#660;&#967;/ &#8216;crazy&#8217; is in bold).</p>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(1)</td>
<td>x-i-b&#8217;e-ki-<bold>ch&#8217;uj</bold>-ir-isa-j</td>
</tr>
<tr>
<td>&#160;</td>
<td><sc>ASP</sc>-1<sc>SG.ABS-DIR</sc>-3<sc>PL.ERG</sc>-crazy-<sc>INCH-CAUS-TRANS</sc></td>
</tr>
<tr>
<td>&#160;</td>
<td>they went somewhere to drive me crazy</td>
</tr>
</tbody>
</table>
</table-wrap>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(2)</td>
<td>qa-<bold>ch&#8217;uj</bold>-ir-isa-x-ik</td>
</tr>
<tr>
<td>&#160;</td>
<td>1<sc>PL.ERG</sc>-crazy-<sc>INCH-CAUS-PASS-NOM</sc></td>
</tr>
<tr>
<td>&#160;</td>
<td>our being driven crazy</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>All modern Mayan writing systems are alphabetic in nature. Literacy in Kaqchikel is currently quite low, in part because written materials are not widely available (on the history of literature and literacy in Kaqchikel, see <xref ref-type="bibr" rid="B124">Maxwell &amp; Hill, 2010</xref>; on literacy in Mayan languages more generally, see <xref ref-type="bibr" rid="B21">Brody, 2004</xref>; <xref ref-type="bibr" rid="B53">England, 1996</xref>, <xref ref-type="bibr" rid="B55">2003</xref>, and references there). While standard orthographies exist for most Mayan languages (<xref ref-type="bibr" rid="B12">Bennett, Coon, &amp; Henderson, 2016</xref>; <xref ref-type="bibr" rid="B101">Kaufman, 2003</xref>), there is a substantial amount of variability in writing conventions across speakers (<xref ref-type="bibr" rid="B26">R. M. Brown et al., 2010</xref>; <xref ref-type="bibr" rid="B53">England, 1996</xref>, <xref ref-type="bibr" rid="B55">2003</xref>). This variation reflects, at least in part, the extensive dialect variation found for those Mayan languages which (like Kaqchikel) are spoken over a wide geographical area (e.g., <xref ref-type="bibr" rid="B26">R. M. Brown et al., 2010</xref>; <xref ref-type="bibr" rid="B122">Majzul et al., 2000</xref>; <xref ref-type="bibr" rid="B124">Maxwell &amp; Hill, 2010</xref>; <xref ref-type="bibr" rid="B155">Richards, 2003</xref>).</p>
</sec>
<sec>
<title>3. Perception study: AX task</title>
<p>To investigate the role that lexical and acoustic experience play in speech perception in Kaqchikel, we carried out a simple AX discrimination task investigating perceptual similarity among stop consonants.</p>
<sec sec-type="methods">
<title>3.1. Method</title>
<p>In this study, Kaqchikel speakers listened to pairs of [CV] or [VC] syllables over headphones. We will sometimes refer to the [CV] condition as the &#8216;Onset&#8217; condition, and the [VC] condition as the &#8216;Coda&#8217; condition. The vowels in a given syllable pair were always identical, but the consonants could be either identical or different. Upon listening to each pair of syllables, the participants were asked to respond Same or Different on a button box. Our underlying assumption is that incorrect Same responses to syllables containing different consonants indicates perceptual similarity between [C<sub>A</sub>]&#126;[C<sub>B</sub>] pairs. Further details of the methodology are outlined below.</p>
<sec>
<title>3.1.1. Participants</title>
<p>Forty-five experimental participants were recruited in Patzic&#237;a, Guatemala (Figure <xref ref-type="fig" rid="F1">1</xref>) by one of the authors (Ajsivinac), who is himself a native speaker of the Patzic&#237;a variety of Kaqchikel. These participants all have self-reported native-level fluency in Kaqchikel, a fact further confirmed by co-author Ajsivinac during conversations before and after the experimental sessions. As is typically the case in Guatemala, most of these participants were also fully bilingual in Spanish. Kaqchikel is nonetheless the primary medium of communication in Patzic&#237;a, and the language most likely spoken by our participants at home and in many public contexts. All communication before, during, and after experimental sessions was conducted in Kaqchikel (the first author, Bennett, is a second-language speaker of Kaqchikel with conversational abilities).</p>
<p>Participants completed a consent form and were given 200 Guatemalan quetzals (&#8776;$27.25) for their participation. All participants were born in the department of Chimaltenango (Figure <xref ref-type="fig" rid="F1">1</xref>), where they also resided at the time of the study. Forty-one participants were born in the town of Patzic&#237;a itself, and 43 were living there at the time of the study. Ages ranged from 18 to 79 years old (Mean: 29, <italic>SD</italic>: 12.3). Thirteen male and 32 female speakers participated in the study (M:F ratio: 0.41). The skew toward female participants is typical of fieldwork in Guatemala, as women typically have greater flexibilty during the workday than men. One participant was excluded from analysis for failure to complete the study.</p>
<p>All experimental sessions were carried out in Patzic&#237;a, Guatemala (Figure <xref ref-type="fig" rid="F1">1</xref>), in a quiet room made available to the authors for the purposes of the study. Each session took about 35 minutes to complete.</p>
</sec>
<sec>
<title>3.1.2. Stimulus design</title>
<p><bold>Recording and pre-processing</bold> The stimuli used in this study were recorded by a male native speaker of Patzic&#237;a Kaqchikel (co-author Ajsivinac). The stimuli were recorded in [pVC] and [CVp] frames. These frames were chosen for several reasons. First, the dominant shape of root morphemes in Kaqchikel (as in other Mayan languages) is /CVC/: There are few content words of the shape /CV/ or /VC/ (e.g., <xref ref-type="bibr" rid="B10">Bennett, 2016</xref> and references there). Recording the materials as [pVC]/[CVp] helps minimize any phonetic artifacts which might result from recording materials that are not native-like in form. Furthermore, /VC/ roots are subject to consonant epenthesis in Kaqchikel, being realized as [&#660;VC] in isolation, which makes it effectively impossible to record simple [VC] syllables. A plain consonant (/p/) was chosen as the frame consonant, rather than an ejective or implosive, to avoid any coarticulatory glottalization on the vowel (see <xref ref-type="bibr" rid="B10">Bennett, 2016</xref>; <xref ref-type="bibr" rid="B13">Bennett et al., in preparation</xref>). The frames [pVC] and [CVp] were recorded for all combinations of the vowels /a i u/ matched with each of the 22 phonemic consonants of Kaqchikel (Table <xref ref-type="table" rid="T1">1</xref>).</p>
<p>The stimuli were presented for recording in random order, using an HTML platform (<xref ref-type="bibr" rid="B51">El Hattab, 2016</xref>). For each stimulus, the speaker was asked to produce 3 repetitions with roughly even intonation. Only the best repetition for each stimulus was selected for further processing and presentation. Each recording was first manually annotated on the segmental level using the acoustic analysis program Praat (<xref ref-type="bibr" rid="B18">Boersma &amp; Weenink, 2016</xref>). A new set of stimuli was then extracted at these segmental boundaries, with the frame consonant /p/ excluded. The exclusion of /p/ was determined on the basis of the waveform, spectrogram, and listener audition (by co-author Bennett). Following Cutler, Weber, Smits, and Cooper (<xref ref-type="bibr" rid="B39">2004</xref>), the stimuli were then amplitude-normalized with respect to the rms amplitude of the vowel (set at 60 dB).</p>
<p><bold>Embedding in noise</bold> In order to increase the likelihood of response errors in our study, we masked the stimuli in speech-shaped noise at a signal-to-noise ratio (SNR) of 0 dB. After amplitude normalization, each stimulus was padded with 250 ms of preceding silence, and 250 ms of following silence. The padded stimuli were then embedded in speech-shaped noise at 0 dB SNR (on our choice of SNR, see <xref ref-type="bibr" rid="B131">Meyer, Dentel, &amp; Meunier, 2013</xref>, <xref ref-type="bibr" rid="B170">Tang, 2015</xref>, Ch. 3.7).<xref ref-type="fn" rid="n2">2</xref></p>
<p><bold>Stimulus pairs</bold> As noted above, participants in this study listened to [CV] or [VC] syllables presented in pairs: [C<sub>A</sub>V]&#126;[C<sub>B</sub>V] or [VC<sub>A</sub>]&#126;[VC<sub>B</sub>]. The perception study was designed to focus on perceptual confusion between the stops /p t k q &#595; t<sup>&#660;</sup> k<sup>&#660;</sup> q<sup>&#660;</sup> &#660;/. Our Target Pairs were pairs of [CV] or [VC] syllables in which both consonants belonged to this set of stops. All other consonants of Kaqchikel were included as fillers in this study, so that participants also heard many filler pairs in which at least one consonant was not a stop. In each [CV]/[VC] syllable the vowel was always one of /a i u/, and vowel quality was always matched between syllables presented in a pair.</p>
<p>There were 270 distinct target pairs in our study, ignoring the order of presentation of the items in each pair. This included 54 Same target pairs (9 stops &#215; 3 vowels &#215; 2 syllable templates) and 216 Different target pairs (<sub>9</sub>C<sub>2</sub> (=36) consonant pairs &#215; 3 vowels &#215; 2 syllable templates). There were 1248 additional filler pairs, which reflect all possible combinations of non-stop consonants (n = 13) with other consonants (n = 22), across 3 vowel and 2 syllable contexts. The ratio of same:different trials in the study was set at 3:4 (including both filler and target pairs).</p>
<p>In order to keep the experiment to a reasonable length, we divided our target pairs into 30 different lists. For each list, we randomly sampled 72 Different target pairs, and 54 Same target pairs. Sampling of Same trials was done with replacement, so that each list could contain multiple instances of a given Same pair (up to a maximum of 3 repetitions).</p>
<p>Within each list we also included 74 filler items composed of consonant pairings that included at least one non-stop consonant. These 74 filler items were sampled with the same 3:4 ratio used to balance same:different trials for the target pairs (32 Same fillers, 42 Different fillers). This resulted in 200 trials per list. The 45 participants were assigned a list in order: Since there were only 30 stimulus lists, the first 15 lists were assigned to two participants each, and the remaining 15 lists assigned to just one participant each. The order of presentation for the pairs in each list was randomized across participants.</p>
</sec>
<sec>
<title>3.1.3. Stimulus presentation</title>
<p>Presentation of the stimuli and logging of participant responses was carried out with a script written in PsychoPy (Version 1.82.01; <xref ref-type="bibr" rid="B139">Peirce, 2007</xref>) and excecuted on a laptop computer. As noted above, this script assigned each participant to one of 30 stimulus lists, and automatically randomized presentation of stimuli within each list.</p>
<p>Prior to the beginning of the experiment, participants were told that they would be listening to a series of syllable pairs, and that they would have to respond as to whether they thought the syllables in each pair were the same or different. They were also told that the stimuli would be embedded in noise of some kind, making them difficult to hear, and that they should not expect the syllables to correspond to actual words of Kaqchikel in most cases (see Section 5.3). This information was provided because pilot testing suggested that the presence of noise and the nonce-word status of the stimuli might be confusing to some participants.</p>
<p>On each trial, participants were first presented with a cross in the center of the screen, lasting 500 ms. The screen then changed to a display showing a green box on the left side of the screen (corresponding to Same responses) and a red box on the right (corresponding to Different responses). Simultaneous with this change in the display, the first member of the stimulus pair for that trial began to play over headphones (Shure SRH 440 over-ear headphones, connected to the computer via an external FiiO E10 USB preamp set at a fixed level across sessions). The order of presentation of the two syllables in a stimulus pair was randomized on each trial.</p>
<p>Upon hearing each stimulus pair, participants responded as to whether they thought the two syllables were identical or different, using a PST Serial Response Box attached to the laptop. Same responses were entered with the leftmost key, and Different responses with the rightmost key; the position of the response keys was not counter-balanced across participants.</p>
<p>Participants were instructed to respond as quickly and as accurately as possible. Participants could take as long as they liked to respond, but trials taking longer than 10 seconds were followed with a reminder to respond as quickly as possible (a yellow warning sign symbol). Even without significant time pressure, participants responded in under one second on most trials (mean RT = 854 ms, median RT = 664 ms). Ten practice trials were completed prior to the actual experiment; these practice trials always involved syllable pairs which were not included in the test list for that speaker&#8217;s session.</p>
<p>After each participant was comfortable with the practice items, they began the actual experiment. The inter-stimulus interval (ISI) between the two stimuli in each pair was set to 300 ms. The inter-trial interval was set at 1500 ms (1000 ms of blank screen followed by the 500 ms cross fixation at the beginning of each trial). Participants were permitted to take a break after every 40 trials. The stimuli were presented at a fixed volume across trials, set at a comfortable level for each listener.</p>
<p>To verify that participants had completed the task as requested, we computed d&#8242; scores (<xref ref-type="bibr" rid="B118">Macmillan &amp; Creelman, 2005</xref>) for perceptual confusions between each pair of stop consonants, collapsing comparisons across all participants. d&#8242; is related to accuracy, but controls for response bias, in particular the tendency to favor one of the two responses regardless of what the stimulus is. The mean d&#8242; score for comparisons in the Onset condition was 1.62, and the mean d&#8242; score for comparisons in the Coda condition was 1.82.</p>
<p>We believe that these are reasonably good d&#8242; values, given that our participants had limited or no prior experience as experimental participants and were not necessarily accustomed to using a computer.<xref ref-type="fn" rid="n3">3</xref> We conclude that the participants in this study completed the task as requested.<xref ref-type="fn" rid="n4">4</xref></p>
<p>A 9-by-9 plot summarizing the d&#8242; scores for all target stop pairs, collapsed across vowel context and syllable position, is provided as an appendix (Appendix D).</p>
</sec>
<sec>
<title>3.1.4. ISI length and processing mode</title>
<p>The ISI used in this study can be estimated in at least two ways. The shortest estimate would be 300 ms, the length of the silent interval between the two stimuli in a given pair. If we also include the noise padding present at the beginning and end of each stimulus, then the ISI would instead be 800 ms in length (250 ms of noise padding before/after each syllable + 300 ms silence between items). We note these values because the duration of the ISI in AX discrimination and related tasks is known to affect the way in which listeners process auditory stimuli (<xref ref-type="bibr" rid="B3">Babel &amp; Johnson, 2010</xref>; <xref ref-type="bibr" rid="B60">Fox, 1984</xref>; <xref ref-type="bibr" rid="B106">Kingston, 2005a</xref>; <xref ref-type="bibr" rid="B108">Kingston, Levy, Rysling, &amp; Staub, 2016</xref>; <xref ref-type="bibr" rid="B130">McGuire, 2010</xref>; <xref ref-type="bibr" rid="B145">Pisoni, 1973</xref>, <xref ref-type="bibr" rid="B146">1975</xref>; <xref ref-type="bibr" rid="B147">Pisoni &amp; Tash, 1974</xref>; <xref ref-type="bibr" rid="B184">Werker &amp; Logan, 1985</xref>; <xref ref-type="bibr" rid="B186">Werker &amp; Tees, 1984b</xref>). Shorter ISIs, particularly those without intervening noise, tend to favor a more acoustically-oriented mode of speech processing which does not necessarily engage the phonemic and lexical levels of speech encoding (i.e., short ISIs encourage a &#8216;prelinguistic&#8217; mode of listening; see Section 8). Longer ISIs, typically at 500 ms or above, seem to condition responses which are more strongly affected by the phonemic and lexical structure of the listener&#8217;s native language (i.e., a &#8216;linguistic&#8217; mode of speech processing). We mention these considerations because an ISI of 800 ms may have facilitated a linguistically-oriented mode of listening, a fact which is relevant given our goal of linking perceptual confusions to statistical facts about words and segments in Kaqchikel.</p>
</sec>
</sec>
</sec>
<sec>
<title>4. Two corpora for Kaqchikel: Assessing the effect of statistical and acoustic factors on speech perception</title>
<p>The overarching goal of this study was to assess the extent to which prior linguistic experience with Kaqchikel might affect consonant discrimination in a perceptual task. To that end, we examined acoustic, segmental, and word-level factors which could play a role in conditioning consonant confusions. Doing so necessitated the development of two corpora for Kaqchikel, which are described in the following sections.</p>
<sec>
<title>4.1. Spoken Kaqchikel: The Solol&#225; corpus</title>
<sec>
<title>4.1.1. Corpus collection</title>
<p>The Solol&#225; corpus is a collection of audio recordings of spontaneous spoken Kaqchikel. This corpus consists of recordings made by two of the authors in Solol&#225;, Guatemala (Figure <xref ref-type="fig" rid="F1">1</xref>) in 2013 (<xref ref-type="bibr" rid="B11">Bennett &amp; Ajsivinac, in preparation</xref>). Sixteen speakers of the Solol&#225; variety of Kaqchikel contributed to this corpus and shared short, spontaneous narratives of their own choosing for the recording.</p>
<p>Fifteen (out of 16) of the speakers were born in the department of Solol&#225;. The remaining speaker was born in the department of Sacatep&#233;quez, to the east of Solol&#225;. As of 2013, the speakers were all living in the department of Solol&#225;, with six living in the city of Solol&#225;, and ten in other towns. Six of these speakers were male, and 10 female; their ages ranged from 19&#8211;84 years old (mean = 33 years, median = 28 years, <italic>SD</italic> = 15.4). The speakers all had self-reported native-level fluency in Kaqchikel, a fact further confirmed by co-author Ajsivinac during conversations before and after the recording sessions. Most speakers reported using Kaqchikel as the primary language of communication at home.</p>
<p>All speakers were recorded using a headset microphone (Audio-Technica ATM73a) and solid-state portable recorder (Fostex FR-2LE), at a 48 kHz sampling rate with 24 bit quantization. The recordings were subsequently downsampled to 16 kHz for forced alignment and acoustic analysis (see below).</p>
</sec>
<sec>
<title>4.1.2. Corpus processing</title>
<p>In total, the corpus amounts to about 4 hours of recorded speech (&#8776;40,000 word tokens). The entire corpus was transcribed orthographically by one of the authors, a native speaker of Kaqchikel (Ajsivinac). We took a subset of this corpus, consisting of approximately 3.5 minutes of audio per speaker (about 50 minutes in total, consisting of 5218 word tokens and 2754 stop consonant tokens), and annotated it phonetically using forced alignment tools. First, the transcriptions in this subset of the corpus were double-checked by another author (Bennett, a trained phonologist and phonetician, as well as an L2 speaker of Kaqchikel with conversational-level abilities). The orthographic transcriptions for this portion of the corpus were then converted into a surface phonetic transcription with a suite of Python scripts (<ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://www.python.org/">http://www.python.org/</ext-link>) implementing grapheme-to-phoneme conversion and several major allophonic rules (see <xref ref-type="bibr" rid="B46">DiCanio et al., 2013</xref> for discussion).<xref ref-type="fn" rid="n5">5</xref></p>
<p>These phonetic transcriptions, and their associated audio, were then submitted to segment-level forced alignment using the Prosodylab-Aligner (<ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://prosodylab.org/tools/aligner/">http://prosodylab.org/tools/aligner/</ext-link>; <xref ref-type="bibr" rid="B75">Gorman, Howell, &amp; Wagner, 2011</xref>). Forced alignment is a computational technique for semi-automatically time-aligning audio files with a corresponding transcription. The Prosodylab-Aligner takes as its input an audio file with an associated sentence-level transcription, and produces a time-aligned Praat TextGrid with annotations at the word and segment levels. An alignment model was first trained on the 50 minute sub-corpus (3 training rounds of 1000 epochs each), then applied to that same corpus to generate the segmental annotations. A total of 2754 stops were annotated at the segmental level using this technique. Alignments were visually-inspected by one of the authors (Bennett, a trained phonologist and phonetician), but not hand-corrected for the purposes of this analysis (see <xref ref-type="bibr" rid="B46">DiCanio et al., 2013</xref> on the distribution of error types in forced alignment).<xref ref-type="fn" rid="n6">6</xref></p>
</sec>
<sec>
<title>4.1.3. Corpus criticism</title>
<p>There are several advantages to using a spoken corpus of this type for acoustic analysis and psycholinguistic research. First, the Solol&#225; corpus is a corpus of spontaneous speech, and is therefore more naturalistic, and more representative of everyday Kaqchikel speech, than a corpus of read or elicited materials. Second, the materials in such a corpus&#8212;which include stories and folktales that are traditionally told in the Solol&#225; region&#8212;may be of greater interest to the Kaqchikel language community than recordings of isolated wordlists or prompted sentences.</p>
<p>There are also potential drawbacks to using a corpus of this type. While the Solol&#225; corpus has the advantage of being naturalistic and thus more ecologically valid than certain other types of audio corpora, the content of the recordings is not controlled in any way (see <xref ref-type="bibr" rid="B191">Xu, 2010</xref> for discussion). As a consequence, data sparsity issues emerge with respect to certain phonetic and phonological structures. For example, ejective /t<sup>&#660;</sup>/ is quite rare in our data (&lt;1% of stops). (This is to be expected, as ejective /t<sup>&#660;</sup>/ is known to have low type and token frequencies in Mayan languages; <xref ref-type="bibr" rid="B10">Bennett, 2016</xref>; <xref ref-type="bibr" rid="B54">England, 2001</xref>.) The paucity of /t<sup>&#660;</sup>/ tokens in the corpus clearly precludes any strong conclusions about the properties of this sound. Additionally, there are relatively few glottalized stops in non-prevocalic (&#8776;coda) position in our corpus (&lt;5% of all stops). This owes in part to the fact that most stops in our corpus, regardless of laryngeal state, occur in pre-vocalic position (85%).</p>
<p>Although the Solol&#225; corpus is a corpus of spontaneous speech, it is also a corpus of monologues rather than dialogues. As such, the speech genre of the corpus may be less than fully naturalistic, and may further show the effects of stereotyped or ritualistic speech patterns associated with storytelling in Kaqchikel. Nonetheless, we believe that the size and composition of this corpus is appropriate for drawing at least some initial conclusions about the phonetic structure of everyday Kaqchikel speech.</p>
</sec>
</sec>
<sec>
<title>4.2. Written corpus</title>
<sec>
<title>4.2.1. Corpus collection</title>
<p>One goal of this paper is to explore how the statistical structure of Kaqchikel&#8212;both in the lexicon (i.e., the vocabulary), and in actual spoken or written usage&#8212;might influence speech perception. To answer this question we needed a reasonably large corpus of written Kaqchikel over which segmental and word-level statistics could be calculated. To the best of our knowledge there are no structured corpora of written Kaqchikel currently available (apart from dictionaries like <xref ref-type="bibr" rid="B116">Macario, Cutzal, &amp; Semey&#225;, 1998</xref>; <xref ref-type="bibr" rid="B121">Majzul, 2007</xref>), and certainly none that are in a digitized, searchable form. It was therefore necessary to construct a novel, digitized written corpus of Kaqchikel in order to assess the statistical patterning of words and segments in the language.</p>
<p>Our corpus is constructed from religious texts, spoken transcripts, government documents, medical handbooks, and other educational books written in Kaqchikel&#8212;essentially all the materials we could find that were already digitized or in an easily digitizable format. The corpus contains approximately 0.7 million word tokens (around 30,000 word types).</p>
<p>The corpus underwent further processing and cleaning before being used to calculate word- and segment-level corpus measures for Kaqchikel. Details on our processing and cleaning methods are given in Appendix A.</p>
</sec>
<sec>
<title>4.2.2. Corpus criticism</title>
<p><bold>Corpus size and composition</bold> Modern corpora of majority languages like English are quite large, on the order of hundreds of millions of word tokens in size (e.g., the Subtlex-UK corpus, 201 million words, <xref ref-type="bibr" rid="B177">van Heuven, Mandera, Keuleers, &amp; Brysbaert, 2014</xref>). Spoken corpora tend to be smaller, but still typically contain several million words (e.g., the Corpus of Spontaneous Japanese, 7 million words, <xref ref-type="bibr" rid="B120">Maekawa, 2003</xref>). Developing corpora of this size is simply not feasible for under-resourced languages like Kaqchikel, which may lack large quantities of written text (particularly digitized text), as well as the economic infrastructure needed to support the collection and annotation of large corpora.</p>
<p>For this reason, in compiling our written corpus we drew on any and all written Kaqchikel texts that we could find. We purposefully excluded dictionaries and collections of neologisms from the corpus because these sources are likely to contain words which are not familiar to most Kaqchikel speakers.</p>
<p>In several respects, our written corpus is far from ideal. First, the corpus is relatively small, containing only &#8776;0.7 million word tokens. It has been argued that a corpus of 16 million word tokens or more is needed for calculating stable estimates of the statistical properties of low frequency words (<xref ref-type="bibr" rid="B28">Brysbaert &amp; New, 2009</xref>).</p>
<p>Second, our corpus contains a mix of both spoken transcripts and written sources. Ideally we would make use of a corpus consisting exclusively of spoken transcripts, given that most Kaqchikel speakers are not literate in the language, or otherwise have limited experience reading in Kaqchikel. Even for majority languages with higher literacy rates, it has been argued that spoken corpora are more representative of speakers&#8217; actual linguistic experience than written corpora (<xref ref-type="bibr" rid="B28">Brysbaert &amp; New, 2009</xref>; <xref ref-type="bibr" rid="B103">Keuleers, Lacey, Rastle, &amp; Brysbaert, 2012</xref>).</p>
<p>Additionally, it is important to recognize that the Kaqchikel orthography is only semi-standardized, and orthographic practices vary across dialects and speakers of the language (<xref ref-type="bibr" rid="B26">R. M. Brown et al., 2010, pp. 3&#8211;4</xref>). For example, our corpus contains both <italic>nb&#8217;&#228;n</italic> and <italic>nub&#8217;&#228;n</italic> as forms of &#8216;(s)he does it,&#8217; reflecting the fact that some dialects omit the 3<sc>SG.ERG</sc> marker -<italic>u</italic>- in particular morphological contexts (<xref ref-type="bibr" rid="B122">Majzul et al., 2000, pp. 69&#8211;70</xref>).</p>
<p><bold>Genre</bold> The <italic>representativeness</italic> of a corpus refers to how closely a corpus reflects actual language use in a particular population (e.g., <xref ref-type="bibr" rid="B1">Atkins, Clear, &amp; Ostler, 1992</xref>; <xref ref-type="bibr" rid="B16">Biber, 1993</xref>). One measure of representativeness is the extent to which the texts and genres included in a corpus correspond to the kinds of texts (or linguistic interactions) that speakers in the target population typically engage with.</p>
<p>The written Kaqchikel corpus described here is not balanced by genre, nor is it particularly speech-like with respect to the thematic content of the materials that it includes. To get a rough sense of how far the corpus deviates from naturalistic speech, we used the transitional probabilities between words in the corpus to create a trigram Markov-chain language model. Using this Markov-chain language model, we stochastically generated (&#8216;babbled&#8217;) some random samples of Kaqchikel. One such sample is shown below.</p>
<p><italic>A sample of Markov-chain Kaqchikel</italic></p>
<disp-quote>
<p>ri taq Mechanpomal moloj: achoq pa ruwi&#8217; yesam&#228;j. K&#8217;&#239;y mul nqak&#8217;axaj nkib&#8217;ij chi ri xaqixaq nuqasaj ri k&#8217;at&#228;n jub&#8217;a&#8217;. K&#8217;o b&#8217;ey chuqa&#8217; nq&#8217;axon nchulun o taq nsinan. K&#8217;&#239;y b&#8217;ey man ntane&#8217; ta ri retal nuya&#8217; chi ke ronojel ri qamolojri&#8217;&#239;l. Richin nawetamaj m&#225;s Rep&#250;blica Democr&#225;tica del Congo Ruanda Jun peraj chi re ri raq&#228;n ya&#8217; Jord&#225;n. Ri Jehov&#225; rik&#8217;in ri m&#225;s &#252;tz chuqa&#8217; man &#252;tz ta yojch&#8217;on rik&#8217;in jun win&#228;q&#8230;</p>
</disp-quote>
<p><italic>Loose English translation of the Markov-chain Kaqchikel sample</italic></p>
<disp-quote>
<p>the Mechanpomal group: on top of what do they work. Many times we listen to what they say about wormwood which lowers the heat a little. There are times too it hurts when he urinates or has sexual relations. Many times it doesn&#8217;t stop, the sign it gives to all of our organizations. In order for you to know more Democratic Republic of the Congo Rwanda A shawl for the river of Jordan. Jehova is with the best and it isn&#8217;t good that we talk to a man&#8230;</p>
</disp-quote>
<p>It is clear from the sample that the written corpus is not particularly speech-like, although it does contain a good range of lexical items covering the topics of religion, geography, and agriculture. Despite the fact that the genres represented by our corpus diverge somewhat from everyday speech, Tang, Bennett, and Ajsivinac (<xref ref-type="bibr" rid="B172">in preparation</xref>) show that word frequencies estimated from this written corpus can be used to predict the duration of words in our corpus of spontaneous spoken Kaqchikel (Section 4.1; see <xref ref-type="bibr" rid="B188">C. E. Wright, 1979</xref> for the classic finding that word frequency and word duration are correlated in English). We take this result as indirect evidence that our written corpus roughly approximates the lexical structure of Kaqchikel as it is actually spoken.</p>
<p>For present purposes, the question is whether measures like functional load or phoneme frequency can be reliably estimated from this corpus. Work in progress (<xref ref-type="bibr" rid="B171">Tang, Bennett, &amp; Ajsivinac, 2015</xref>) suggests that estimates of these measures are stable even over small sub-samples of this corpus (e.g., 20,000 words; see too <xref ref-type="bibr" rid="B47">Dockum &amp; Campbell-Taylor, 2017</xref>; <xref ref-type="bibr" rid="B71">Gasser &amp; Bowern, 2014</xref>; <xref ref-type="bibr" rid="B117">Macklin-Cordes &amp; Round, 2015</xref>). As such, we believe that our corpus is indeed of sufficient size for the estimation of these measures.</p>
</sec>
</sec>
</sec>
<sec>
<title>5. Predictions and model design</title>
<p>In this section we consider how acoustic factors, word-level statistics, and segmental statistics might interact with speech perception in Kaqchikel. We also describe the basic modeling procedure we used to test whether these predictions were borne out in our results (Section 6).</p>
<sec>
<title>5.1. Statistical model</title>
<p>We analyzed participant accuracy on each trial of the AX discrimination task with a mixed effects logistic regression in R (<xref ref-type="bibr" rid="B151">R Development Core Team, 2013</xref>), using the glmer function in the lme4 library (<xref ref-type="bibr" rid="B7">Bates, Maechler, &amp; Bolker 2011</xref>). Recall that each trial could either contain two identical stimuli (the Same condition), or two different stimuli (the Different condition). We interpreted incorrect responses on Different trials as evidence that a given pair of stimuli was perceptually similar (having been mistaken as identical). Same trials are not similarly informative regarding perceptual confusion; we therefore analyzed only the accuracy of participant response on Different trials.<xref ref-type="fn" rid="n7">7</xref></p>
</sec>
<sec>
<title>5.2. Acoustic similarity</title>
<p>One of our main expectations is that greater acoustic similarity between a pair of syllables should predict greater perceptual similarity between those syllables (e.g., <xref ref-type="bibr" rid="B48">Dubno &amp; Levitt, 1981</xref>). Two acoustic similarity measures were considered. The first measure is Stimulus Similarity&#8212;the raw acoustic similarity of the stimuli themselves. We expected stimulus similarity to have a substantial effect on stimulus discrimination in our study.</p>
<p>The second measure is Category Similarity&#8212;the similarity of two phoneme categories based on <italic>prior phonetic experience</italic>. Following a large body of work in Exemplar Theory, we assume that phonemic categories are associated with episodic memory traces (or exemplars), which are phonetically-rich representations of that category as previously encountered on specific occasions in actual speech (<xref ref-type="bibr" rid="B63">Gahl &amp; Yu, 2006</xref>; <xref ref-type="bibr" rid="B72">Goldinger, 1996</xref>, <xref ref-type="bibr" rid="B73">1998</xref>; <xref ref-type="bibr" rid="B141">Pierrehumbert, 2001</xref>, <xref ref-type="bibr" rid="B142">2002</xref>; <xref ref-type="bibr" rid="B181">Wedel, 2004</xref> and references there). On this view, the category similarity of two phonemes can be conceived of as the extent to which their exemplar clouds show overlap in phonetic space (see also <xref ref-type="bibr" rid="B194">Yu, 2011</xref>).</p>
<p>Figures <xref ref-type="fig" rid="F2">2</xref> and <xref ref-type="fig" rid="F3">3</xref> illustrate the importance of distinguishing stimulus similarity and category similarity. These figures show two hypothetical exemplar clouds for the phonemes /k/ and /p/ over some acoustic dimension(s) (say, VOT and burst intensity). The two clouds represent a collection of prior phonetic experiences that the listener has associated with each phoneme. In the context of our study, the two dots represent two stimuli presented to the listener for discrimination (say [ka] and [pa]).</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Non-overlapping exemplar clouds.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6226/file/75171/"/>
</fig>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Overlapping exemplar clouds.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6226/file/75172/"/>
</fig>
<p>In Figure <xref ref-type="fig" rid="F2">2</xref>, the exemplar clouds are non-overlapping (i.e., VOT values for /k/ and /p/ are typically quite distinct). As a consequence, listeners would likely conclude that the two stimuli (the dots) belong to different categories. In Figure <xref ref-type="fig" rid="F3">3</xref>, the clouds are substantially more overlapped, with the two stimuli falling in the overlapping region. This overlap between the two categories increases the level of uncertainty for the listener, making it more difficult to determine whether the two stimuli belong to different phonemic categories. To the extent that overlap along particular phonetic dimensions makes listeners less likely to rely on those dimensions for category discrimination (e.g., <xref ref-type="bibr" rid="B92">Holt &amp; Lotto, 2006</xref>), category overlap may influence discrimination even for stimuli which are acoustically unambiguous (i.e., in non-overlapping regions of Figure <xref ref-type="fig" rid="F3">3</xref>), by reducing overall sensitivity to certain potential cues to consonant identity.</p>
<sec>
<title>5.2.1. Stimulus similarity</title>
<p>It is intuitively clear that higher levels of acoustic similarity between stimuli should lead to higher rates of confusion between those stimuli in a discrimination task (e.g., <xref ref-type="bibr" rid="B48">Dubno &amp; Levitt, 1981</xref>; <xref ref-type="bibr" rid="B81">Hall &amp; Hume, in preparation</xref>; <xref ref-type="bibr" rid="B152">Redford &amp; Diehl, 1999</xref> and many others). To evaluate the importance of other factors in this study&#8212;particularly those related to prior phonetic experience (category similarity)&#8212;stimulus similarity must therefore be included as a control predictor.</p>
<p>To capture the acoustic similarity between the stimulus pairs in each trial, an acoustic distance metric was applied to each stimulus pair (after embedding the stimuli in noise). Such a metric should allow us to capture the raw acoustic information that could be used by the listeners to perform the AX task, even without accessing higher-level perceptual, phonemic, or lexical processing. Our acoustic distance metric was calculated using Phonological CorpusTools (<xref ref-type="bibr" rid="B80">Hall, Allen, Fry, Mackie, &amp; McAuliffe, 2015</xref>). First, the waveform of each stimulus was transformed into mel-frequency cepstrum coefficents (MFCCs) (<xref ref-type="bibr" rid="B132">Mielke, 2012</xref>), a common re-representation of the acoustic signal used widely in speech recognition research. The number of MFCCs was set to 12, as this allows the model to capture speaker-independent information about acoustic similarity (see <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://corpustools.readthedocs.io/en/latest/acoustic_similarity.html">http://corpustools.readthedocs.io/en/latest/acoustic_similarity.html</ext-link>). Dynamic Time Warping (DTW), another common speech processing technique, was used to compute an explicit distance metric on the basis of the MFCC-transformed stimuli (<xref ref-type="bibr" rid="B132">Mielke, 2012</xref>; <xref ref-type="bibr" rid="B158">Sakoe &amp; Chiba, 1971</xref>).</p>
<p>Our statistical model included a fixed effect for Stimulus Similarity, representing the acoustic distance between two stimuli according to this DTW metric. This predictor was <italic>z</italic>-score normalized using the <monospace>scale()</monospace> function in R.</p>
</sec>
<sec>
<title>5.2.2. Category similarity</title>
<p>To calculate the acoustic similarity of two stops at the phonemic level (category similarity), we computed DTW over pairs of stops as they occur in our acoustic corpus (Section 4.1). We further limited our comparisons to stop consonants occuring in similar environments, since the confusability of any given pair of stops may depend on the phonotactic and prosodic context (e.g., <xref ref-type="bibr" rid="B34">Chang, Plauch&#233;, &amp; Ohala, 2001</xref>; <xref ref-type="bibr" rid="B39">Cutler et al., 2004</xref>).</p>
<list list-type="order">
<list-item><p>Using our acoustic corpus, we identified all instances of /p t k q &#595; t<sup>&#660;</sup> k<sup>&#660;</sup> q<sup>&#660;</sup>/. These were divided into two groups: (a) pre-vocalic ([<underline>C</underline>V]); and (b) post-vocalic, but non-prevocalic ([V<underline>C</underline>(C/#)]).</p></list-item>
<list-item><p>Using the segmentation provided by forced alignment (Section 4.1), the waveforms corresponding to each target stop consonant and the vowel adjacent to it were extracted individually. This gave two sets of waveforms corresponding to /p t k q &#595; t<sup>&#660;</sup> k<sup>&#660;</sup> q<sup>&#660;</sup>/ in [CV] and [VC] transitions.</p></list-item>
<list-item><p>These waveforms were further divided into subsets on the basis of the vowel, matching [CV] and [VC] waveforms according to the quality and stress profile of the vowel. For example, [&#39;ke] could be compared with [&#39;te], but not with [te], [&#39;ti], [&#39;k&#603;], or [&#39;ek].</p></list-item>
<list-item><p>Within each matched [CV] or [VC] subset, we computed an acoustic distance measure (DTW) between all pairs of waveforms within that set which contained different stop consonants. For example, if [&#39;ke] occurred twice in the corpus, and [&#39;te] occurred three times, we would compute six pairwise acoustic distance measures: [&#39;ke]<sub>1</sub>&#126;[&#39;te]<sub>1</sub>, [&#39;ke]<sub>1</sub>&#126;[&#39;te]<sub>2</sub>, [&#39;ke]<sub>1</sub>&#126;[&#39;te]<sub>3</sub>; and [&#39;ke]<sub>2</sub>&#126;[&#39;te]<sub>1</sub>, [&#39;ke]<sub>2</sub>&#126;[&#39;te]<sub>2</sub>, [&#39;ke]<sub>2</sub>&#126;[&#39;te]<sub>3</sub>.</p></list-item>
<list-item><p>The outcome of this procedure is a set of acoustic distances between tokens of the stop categories /p t k q &#595; t<sup>&#660;</sup> k<sup>&#660;</sup> q<sup>&#660;</sup>/, grouped according to their syllabic context (onset/coda) and the properties of the preceding/following vowel. As an aggregate measure of category similarity, we took the mean and standard deviation of these values for each pair of stops.</p></list-item>
</list>
<p>These measures of category similarity were then used as predictors in our analysis of perceptual similarity (Section 6).<xref ref-type="fn" rid="n8">8</xref></p>
<p>Under the assumption that each stop token in the corpus counts as an exemplar, and exemplars are clustered together in clouds according to their category membership, the mean category distance between two stops can be interpreted as the distance between the two centroids of the exemplar clouds associated with each phoneme. The standard deviation of the distances between tokens is a measure of how <italic>consistently</italic> different the two categories are, across contexts and repetitions. These measures are logically and practically independent of each other. For instance, /&#595;/&#126;/p/ and /t&#865;&#643;<sup>&#660;</sup>/&#126;/q<sup>&#660;</sup>/ have similar mean acoustic distances (51.04 and 51.74 respectively), but the standard deviations of the acoustic distances associated with each category are rather different (8.80 and 5.39 respectively). Since the separation between category means and the variance around those means might both matter for the overall separation of two phonemic categories in the acoustic space, we treated both measures of category similarity as predictors in our analysis of the perception study described above. Ultimately, only the category means proved to be a reliable predictor of perceptual similarity in our study (Section 6).</p>
<p>Our statistical model includes two fixed effects for Category Similarity between the two phonemic categories being compared on a given Different trial: Mean Category Similarity and Standard Deviation of Category Similarity. Both predictors were <italic>z</italic>-score normalized using the <monospace>scale()</monospace> function in R.</p>
</sec>
</sec>
<sec>
<title>5.3. Word- and segment-level statistical factors</title>
<p>The analysis of statistical effects on speech perception in Kaqchikel took into account a number of distinct segment- and word-level predictors. Only a few of these predictors made a significant contribution to predicting patterns of perceptual confusion between stops in our study (Section 6). In the following section we define only those predictors which made a reliable contribution to predicting stop discrimination in our study, and leave the definition of the other, non-significant factors which we considered to Appendix C.</p>
<sec>
<title>5.3.1. Segment-level factors</title>
<p>Three segment-level predictors were considered: segmental frequency, functional load, and distributional overlap. Of these, only functional load and distributional overlap emerged as significant predictors of perceptual confusions in our study.</p>
<p><bold>Functional load</bold> Intuitively, Functional Load characterizes the importance of a given phonemic contrast for distinguishing words in a language. One way of defining the Functional Load of two phonemes in a language is to count the number of minimal pairs that are distinguished solely by the contrast between those phonemes (<xref ref-type="bibr" rid="B91">Hockett, 1967</xref>; <xref ref-type="bibr" rid="B111">Ku&#269;era, 1963</xref>; <xref ref-type="bibr" rid="B123">Martinet, 1952</xref>; <xref ref-type="bibr" rid="B167">Surendran &amp; Niyogi, 2003</xref>, <xref ref-type="bibr" rid="B168">2006</xref>). It has been argued that Functional Load and related measures condition the probability of diachronic phoneme mergers (<xref ref-type="bibr" rid="B182">Wedel, Jackson, &amp; Kaplan, 2013</xref>; <xref ref-type="bibr" rid="B183">Wedel, Kaplan, &amp; Jackson, 2013</xref>), as well as the production of phonemic contrasts (<xref ref-type="bibr" rid="B4">Baese-Berk &amp; Goldrick, 2009</xref>; <xref ref-type="bibr" rid="B74">Goldrick, Vaughn, &amp; Murphy, 2013</xref>; <xref ref-type="bibr" rid="B134">Nelson &amp; Wedel, 2017</xref>).</p>
<p>Perhaps most relevant to this study, Functional Load may also interact with the <italic>perception</italic> of phonemic contrasts. Graff (<xref ref-type="bibr" rid="B76">2012</xref>), drawing on data from 60 languages and 25 language families, argues that languages tend to use perceptually robust phoneme contrasts to distinguish minimal pairs. Consequently, there should be a positive correlation between functional load and the perceptual distinctiveness of a given phonemic contrast. Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) and Stevenson and Zamuner (<xref ref-type="bibr" rid="B166">2017</xref>) show that in French, vowel pairs which have a higher functional load are also more perceptually distinct, even when other factors (such as raw acoustic similarity) are taken into account (see also <xref ref-type="bibr" rid="B153">Renwick, 2014</xref> for similar claims about Romanian). Relatedly, L. Davidson, Shaw, and Adams (<xref ref-type="bibr" rid="B43">2007</xref>) found that listeners were more attentive to subtle phonetic details in an AX discrimination task (such as the presence vs. absence of schwa in clusters, [C&#601;C]&#126;[CC]) when the items constituted minimal pairs.</p>
<p>For this study, we focused on a metric of functional load which captures the change in entropy of the lexicon following merger of a phoneme contrast. This metric is sometimes called <italic>lexical</italic> &#916;-<italic>entropy</italic> (for comparison with other metrics, see footnote 12). To compute this measure, we employed an information-theoretic method (<xref ref-type="bibr" rid="B160">Shannon, 1948</xref>). We first calculated the entropy of the Kaqchikel lexicon&#8212;a measure of uncertainty&#8212;using Equation 1 (<xref ref-type="bibr" rid="B167">Surendran &amp; Niyogi, 2003</xref>, <xref ref-type="bibr" rid="B168">2006</xref>). For our purposes, the entropy <italic>H</italic>(<italic>L</italic>) of a language measures how diverse the vocabulary is (basically the size of the lexicon), weighted by token frequency. Entropy depends on <italic>p<sub>w</sub></italic>, the probability of a given word <italic>w</italic> in our written corpus. Functional load (Equation 2) is measured by estimating how much the lexicon &#8216;shrinks&#8217; when two phonemes are merged into one (i.e., the number of distinct words made homophonous, weighted by frequency). The entropy of a lexicon in which phonemes <italic>x, y</italic> have been merged, <italic>H</italic>(<italic>L<sub>xy</sub></italic>), is compared to the original, non-merged lexicon <italic>H</italic>(<italic>L</italic>) to yield the functional load of the <italic>x, y</italic> contrast (Equation 2). Phoneme pairs with a higher functional load should lead to a larger proportional decrease in entropy when they are merged.</p>
<disp-formula id="FD1">
<label>Eq. 1</label>
<alternatives>
<mml:math id="Eq001-mml"><mml:mrow><mml:mi>H</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>L</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>&#x03A3;</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mtext>&#x2009;</mml:mtext><mml:msub><mml:mi>p</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mo>&#x00D7;</mml:mo><mml:mtext mathvariant="italic">lo</mml:mtext><mml:msub><mml:mi>g</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>p</mml:mi><mml:mi>w</mml:mi></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow></mml:math>
<tex-math id="M1">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
H\left(L \right) = - {\Sigma _w}\;{p_w} \times lo{g_2}({p_w})
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75173/"/>
</alternatives>
</disp-formula>
<disp-formula id="FD2">
<label>Eq. 2</label>
<alternatives>
<mml:math id="Eq002-mml"><mml:mrow><mml:mtext mathvariant="italic">FL</mml:mtext><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mi>x</mml:mi><mml:mo>,</mml:mo><mml:mi>y</mml:mi></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>H</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>L</mml:mi><mml:mo>)</mml:mo></mml:mrow><mml:mo>&#x2212;</mml:mo><mml:mi>H</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:msub><mml:mi>L</mml:mi><mml:mrow><mml:mtext mathvariant="italic">xy</mml:mtext></mml:mrow></mml:msub></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow><mml:mrow><mml:mi>H</mml:mi><mml:mrow><mml:mo>(</mml:mo><mml:mi>L</mml:mi><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:mfrac></mml:mrow></mml:math>
<tex-math id="M2">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
FL\left({x,y} \right) = \frac{{H\left(L \right) - H\left({{L_{xy}}} \right)}}{{H\left(L \right)}}
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75174/"/>
</alternatives>
</disp-formula>
<p>For the purpose of computing functional load, &#8216;words&#8217; are defined as whole word forms, including affixes (i.e., as strings of segments separated by white space in a text; see Appendix A).</p>
<p>Wedel, Jackson, and Kaplan (<xref ref-type="bibr" rid="B182">2013</xref>) found that patterns of diachronic phoneme merger were better predicted by a measure of functional load which only compares words belonging to the same lexical category (e.g., two nouns distinguished by an /A B/ phoneme contrast would contribute to the functional load of /A B/, but not a noun-verb pair). As Kaqchikel is moderately agglutinating (Section 2), many words bear affixes which unambiguously indicate their part of speech (e.g., both affixes in <italic>r-utz-il</italic> &#8216;its goodness&#8217; 3<sc>SG.ERG</sc>-good-<sc>NOM</sc> signal that this word is a noun). Our whole-word measure of functional load is thus probably biased toward comparing words within the same lexical category, as in Wedel, Jackson, and Kaplan (<xref ref-type="bibr" rid="B182">2013</xref>). Unlike Wedel, Jackson, and Kaplan (<xref ref-type="bibr" rid="B182">2013</xref>), we did not consider measures of functional load computed over lemmas (basically, uninflected stems) because a lemmatized corpus of Kaqchikel is not currently available.</p>
<p>A fixed effect of Functional Load was included in our statistical model, reflecting the measure of &#916;-entropy described above.</p>
<p><bold>Distributional overlap</bold> Recent work by Hall et al. (<xref ref-type="bibr" rid="B83">2014</xref>) and Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) argues that the predictability of two phonemes across contexts contributes to their perceptual confusability. The theoretical context of this claim is one in which contrastiveness is assumed to be gradient rather than categorical: Two phonemes which occur in many of the same environments are taken to be <italic>more contrastive</italic> (i.e., less predictable) than phonemes which mostly occur in distinct environments (<xref ref-type="bibr" rid="B78">Hall, 2012</xref>, <xref ref-type="bibr" rid="B79">2013</xref>). By hypothesis, phoneme pairs which are more contrastive (less predictable from context) are expected to be more readily discriminated.<xref ref-type="fn" rid="n9">9</xref></p>
<p>We took Jeffreys&#8217; distance (also called Jeffrey divergence; henceforth JD) as our measure of the distributional overlap (=contextual predictability) of two phonemes. JD determines the contextual probability of two phonemes based on the local segmental contexts X__Y that they occur in. JD can thus be interpreted as a measure of phonotactic similarity.<xref ref-type="fn" rid="n10">10</xref></p>
<p>JD was implemented using the <italic>TiMBL</italic> manual (<xref ref-type="bibr" rid="B40">Daelemans, Zavrel, van der Sloot, &amp; van den Bosch, 2009, p. 26</xref>). In our study, each segment type is a class, and our features are the presence and absence of segments; more specifically, we used a trigram sliding window to generate features, with the target segment being in the first, second, or third position of the trigram window. Trigram windows are commonly used to capture phonotactics in computational linguistics (e.g., <xref ref-type="bibr" rid="B98">Jurafsky, Bell, Gregory, &amp; Raymond, 2001</xref>) as well as phonology more broadly (e.g., <xref ref-type="bibr" rid="B88">Hayes &amp; Wilson, 2008</xref>). In the specific case of Kaqchikel, trigram windows are necessary to capture certain co-occurrence restrictions which hold between the two consonants in a /CVC/ root (see <xref ref-type="bibr" rid="B10">Bennett, 2016</xref>; <xref ref-type="bibr" rid="B13">Bennett et al., in preparation</xref>).</p>
<p>A major difference between our metric and other metrics (such as the entropy-based measures in <xref ref-type="bibr" rid="B140">Peperkamp et al., 2006</xref> and <xref ref-type="bibr" rid="B78">Hall, 2012</xref>) is that our contexts are defined over segments as opposed to phonological features. This decision owes in part to our own uncertainty about which phonological features are most appropriate for classifying segments in Kaqchikel, particularly in the case of laryngeal contrasts (see <xref ref-type="bibr" rid="B13">Bennett et al., in preparation for discussion</xref>). The details of the metric are shown below in Equation (3).</p>
<disp-formula id="FD3">
<label>Eq. 3</label>
<alternatives>
<mml:math id="Eq003-mml"><mml:mrow><mml:mtext mathvariant="italic">JD</mml:mtext><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo>,</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mo>=</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mstyle displaystyle='true'><mml:munderover><mml:mo>&#x2211;</mml:mo><mml:mi>i</mml:mi><mml:mi>n</mml:mi></mml:munderover><mml:mi>P</mml:mi></mml:mstyle><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x007C;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mo>&#x00D7;</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mtext mathvariant="italic">log</mml:mtext><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x007C;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mi>m</mml:mi></mml:mfrac></mml:mrow><mml:mo>)</mml:mo></mml:mrow><mml:mtext>&#x2009;</mml:mtext><mml:mo>+</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x007C;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mo>&#x00D7;</mml:mo><mml:mtext>&#x2009;</mml:mtext><mml:mtext mathvariant="italic">log</mml:mtext><mml:mrow><mml:mo>(</mml:mo><mml:mrow><mml:mfrac><mml:mrow><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x007C;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mi>m</mml:mi></mml:mfrac></mml:mrow><mml:mo>)</mml:mo></mml:mrow></mml:mrow></mml:math>
<tex-math id="M3">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
JD({S_1},{S_2}) = \sum\limits_i^n P ({C_i}|{S_1}) \times log\left({\frac{{P({C_i}|{S_1})}}{m}} \right) + P({C_i}|{S_2}) \times log\left({\frac{{P({C_i}|{S_2})}}{m}} \right)
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75175/"/>
</alternatives>
</disp-formula>
<list list-type="bullet">
<list-item><p><inline-formula>
<alternatives>
<mml:math id="Eq004-mml"><mml:mrow><mml:mi>m</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x007C;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>1</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo><mml:mo>+</mml:mo><mml:mi>P</mml:mi><mml:mo stretchy='false'>(</mml:mo><mml:msub><mml:mi>C</mml:mi><mml:mi>i</mml:mi></mml:msub><mml:mo>&#x007C;</mml:mo><mml:msub><mml:mi>S</mml:mi><mml:mn>2</mml:mn></mml:msub><mml:mo stretchy='false'>)</mml:mo></mml:mrow><mml:mn>2</mml:mn></mml:mfrac></mml:mrow></mml:math>
<tex-math id="M4">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
m = \frac{{P({C_i}|{S_1}) + P({C_i}|{S_2})}}{2}
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75176/"/>
</alternatives>
</inline-formula></p></list-item>
<list-item><p><italic>S</italic><sub>1</sub> and <italic>S</italic><sub>2</sub> are two phones.</p></list-item>
<list-item><p><italic>C<sub>i</sub></italic> is a phonotactic environment, defined with a sliding trigram window: __XY, X__Y, XY___</p></list-item>
</list>
<p>A fixed effect of Distributional Overlap was included in our statistical model, reflecting this measure of JD.<xref ref-type="fn" rid="n11">11</xref></p>
</sec>
<sec>
<title>5.3.2. Word-level factors</title>
<p>In addition to the variables mentioned thus far, which were all part of the experimental design, a number of nuisance variables were included in the analysis of consonant confusions: wordhood, word frequency, neighborhood density, average neighborhood frequency, and bigram frequency. Our study was not designed to test the effect of these factors, but we included them in the analysis as control predictors, just in case they had an effect on our results. Of these factors, only word frequency made an appreciable contribution to predicting consonant discrimination in our study (and even then, only marginally so). We describe wordhood and word frequency here; the remaining predictors are defined in Appendix C.</p>
<p><bold>Wordhood</bold> Ganong (<xref ref-type="bibr" rid="B69">1980</xref>) established the now classic result that listeners are more likely to identify a phonetically ambiguous segment as belonging to some phoneme P<sub><italic>x</italic></sub> if the categorization of that segment as P<sub><italic>x</italic></sub> results in an actual word of the listener&#8217;s native language, and categorizing the segment as a competing phoneme P<sub><italic>y</italic></sub> does not (see <xref ref-type="bibr" rid="B60">Fox, 1984</xref>; <xref ref-type="bibr" rid="B106">Kingston, 2005a</xref>; <xref ref-type="bibr" rid="B108">Kingston et al., 2016</xref>; <xref ref-type="bibr" rid="B148">Pitt &amp; Samuel, 1993</xref> and references there).</p>
<p>While our stimuli did contain real words, they were only included in order to achieve balanced coverage over the consonant and vowel combinations which were the focus of this study. However, since wordhood is known to play a role in speech perception, it was included as a possible predictor of perceptual confusions in our AX discrimination task.</p>
<p>To assess the wordhood of our stimuli, we consulted two sources: a native speaker of Kaqchikel (co-author Ajsivinac) and the headwords in two dictionaries (<xref ref-type="bibr" rid="B116">Macario et al., 1998</xref>; <xref ref-type="bibr" rid="B121">Majzul, 2007</xref>). We considered a [CV] or [VC] stimulus to be a &#8216;word&#8217; of Kaqchikel if it was identical to either a function word (e.g., the particle <italic>k&#8217;a</italic> /k<sup>&#660;</sup>a/ &#8216;until, then, well&#8217;) or a content word (e.g., <italic>aq&#8217;</italic> /aq<sup>&#660;</sup>/ &#8594; [&#660;aq<sup>&#660;</sup>] &#8216;tongue&#8217;; word-initial epenthetic glottal stops were ignored for the purposes of computing wordhood). Affixes and other bound morphemes were not considered to be words in this sense (e.g., <italic>at</italic>- /at-/ 2<sc>SG.ABS</sc> or -<italic>i&#8217;</italic> /-i&#660;/ <sc>REFLEXIVE</sc>). We tailored these judgments to the Patzic&#237;a dialect: For example, <italic>uq</italic> [&#660;uq] counted as a word because <italic>&#252;q</italic> &#8216;skirt&#8217; is pronounced as [&#660;uq] (rather than historical [&#660;&#650;q]) in the Patzic&#237;a dialect. Only 15 of our experimental items (including fillers) were actual words of Kaqchikel; the remainder (108) were non-words according to these criteria.</p>
<p>A fixed effect for Wordhood was included in our statistical model as the absolute difference of the wordhood values of two given stimuli: If both stimuli were words or both were non-words, the value was coded as 0, otherwise as 1. This predictor was <italic>z</italic>-score normalized using the <monospace>scale()</monospace> function in R.</p>
<p><bold>Word frequency</bold> The effect of wordhood on phonemic categorization also obtains when the categorization of an ambiguous segment as <italic>either</italic> phoneme P<sub><italic>x</italic></sub> or phoneme P<sub><italic>y</italic></sub> would result in a real word, but the two resulting words differ in token frequency (<xref ref-type="bibr" rid="B44">de Marneffe, Tomlinson Jr., Tice, &amp; Sumner, 2011</xref>). This suggests that categorization judgments can be influenced not only by the categorical word&#126;non-word distinction, but also by gradient differences in word frequency.</p>
<p>More generally, word frequency has shown to contribute to both visual and auditory word recognition, with high frequency words being recognized more accurately and more quickly than low frequency words (<xref ref-type="bibr" rid="B20">Broadbent, 1967</xref>; <xref ref-type="bibr" rid="B25">C. R. Brown and Rubenstein, 1961</xref>; <xref ref-type="bibr" rid="B57">Felty, Buchwald, Gruenenfelder, and Pisoni, 2013</xref>; <xref ref-type="bibr" rid="B93">Howes, 1957</xref>; <xref ref-type="bibr" rid="B170">Tang, 2015</xref>, Ch. 4). Furthermore, when words are incorrectly identified in perception, the perceived word tends to have roughly the same lexical frequency as the intended word (<xref ref-type="bibr" rid="B170">Tang, 2015</xref>, Ch. 4; <xref ref-type="bibr" rid="B173">Tang &amp; Nevins, 2014</xref>; <xref ref-type="bibr" rid="B174">Tang &amp; Nevins, in preparation</xref>; <xref ref-type="bibr" rid="B178">Vitevitch, 2002</xref>). It follows that in our study, even if the stimuli (one or both) were incorrectly perceived on a given trial, the difference in word frequency between the two items could still bias the participants&#8217; responses.</p>
<p>Word frequency was obtained using our written corpus. We only obtained word frequency information if a stimulus was determined to be a word according the criteria described above. We used the difference in frequency between the two stimuli on a given trial as a predictor of consonant confusions; non-words were coded as having zero frequency.</p>
<p>A fixed effect of Word Frequency was included in our statistical model as the absolute difference of the log-transformed (base-10) word token frequencies of the stimuli in each trial, with Laplace (&#8216;add one&#8217;) smoothing for frequencies of zero (prior to log-transformation; <xref ref-type="bibr" rid="B27">Brysbaert &amp; Diependaele, 2013</xref>). This predictor was <italic>z</italic>-score normalized using the <monospace>scale()</monospace> function in R.</p>
</sec>
</sec>
</sec>
<sec>
<title>6. Analysis and results</title>
<sec>
<title>6.1. Statistical modeling</title>
<p>The statistical analysis began with the construction of an initial (or &#8216;superset&#8217;) model which included a large number of predictors. This initial model was then simplified through a model criticism procedure described in Appendix B. The factors included in the initial model are described below.</p>
<sec>
<title>6.1.1. Fixed effects</title>
<p>As mentioned above, our initial model included fixed effects for three acoustic predictors (Stimulus Similarity, Mean Category Similarity, and Standard Deviation of Category Similarity), as well as fixed effects for Functional Load, Distributional Overlap, Wordhood, and Word Frequency. Of these factors, only Stimulus Similarity, Mean Category Similarity, Functional Load, Distributional Overlap, and Word Frequency emerged as significant predictors of consonant discrimination in the final statistical model.</p>
<p>Along with these predictive factors, our initial model included fixed effects for Segmental Frequency, Neighborhood Density, Average Neighborhood Frequency, and Bigram Frequency (see Appendices B and C). These predictors were coded by log-transforming the values of the relevant measure, and taking the absolute difference of those log-transformed values for the two stimuli on each trial (Laplace smoothing was also used for Word Frequency). All five predictors were <italic>z</italic>-score transformed; none of them emerged as predictive in our final statistical model.</p>
<p><bold>Response time</bold> There is a well-known trade-off between speed and accuracy in many behavioral tasks (e.g., <xref ref-type="bibr" rid="B89">Heitz, 2014</xref>). To account for the possibility of such a tradeoff in our study, Response Time was treated as a fixed effect predictor in the analysis of accuracy (<xref ref-type="bibr" rid="B42">D. Davidson &amp; Martin, 2013</xref>).</p>
<p>The response time on each trial was measured from the offset of the second stimulus (including the 250 ms of noise padding following the end of the syllable itself) to the time at which the response was logged. Given that each participant might have a different baseline response speed, these response times were transformed into by-participant <italic>z</italic>-scores.</p>
</sec>
<sec>
<title>6.1.2. Random effects</title>
<p><bold>Item-level random effects</bold> Unordered Stimulus Pair was treated as a random intercept, and was defined as the unordered pairing of any two Different stimuli. As each participant heard one of 30 different lists of stimulus pairs (Section 3.1.2), List was also included as a random intercept. The order of the two stimuli in a given trial (Stimulus Order) was included as another random intercept, since the order of stimulus presentation has been reported to affect same-different discrimination judgments in some tasks (<xref ref-type="bibr" rid="B15">Best et al., 2001</xref>; <xref ref-type="bibr" rid="B30">Bundgaard-Nielsen, Baker, Kroos, Harvey, &amp; Best, 2015</xref>; <xref ref-type="bibr" rid="B37">Cowan &amp; Morse, 1986</xref>; <xref ref-type="bibr" rid="B41">Dar, Keren-Portnoy, &amp; Vihman, 2018</xref>; <xref ref-type="bibr" rid="B154">Repp &amp; Crowder, 1990</xref>).</p>
<p>Finally, the position of the target stop in each stimulus (Onset vs. Coda) was treated as a random intercept. Phonotactic context is an important factor that influences consonant discrimination (see <xref ref-type="bibr" rid="B189">R. Wright, 2004</xref> for an overview). A large body of research has found that place, manner, and laryngeal features are better discriminated for prevocalic [CV] consonants (particularly stops) than for non-prevocalic [VC] consonants (e.g., <xref ref-type="bibr" rid="B8">Benk&#237;, 2003</xref>; <xref ref-type="bibr" rid="B17">Bladon, 1986</xref>; <xref ref-type="bibr" rid="B62">Fujimura, Macchi, &amp; Streeter, 1978</xref>; <xref ref-type="bibr" rid="B97">Jun, 2004</xref>; <xref ref-type="bibr" rid="B152">Redford &amp; Diehl, 1999</xref>; <xref ref-type="bibr" rid="B164">Steriade, 2001</xref>, <xref ref-type="bibr" rid="B165">2009</xref>; <xref ref-type="bibr" rid="B170">Tang, 2015</xref>; <xref ref-type="bibr" rid="B180">Wang &amp; Bilger, 1973</xref>, and others; cf. <xref ref-type="bibr" rid="B39">Cutler et al., 2004</xref>; <xref ref-type="bibr" rid="B131">Meyer et al., 2013</xref> for skeptical views). Our initial analysis of d&#8242; found no distinction in perceptibility between onset [CV] and coda [VC] contrasts (Section 2), but it still seemed prudent to include consonant position as a potential predictor of consonant confusions in this study.</p>
<p><bold>Participant-level random effects</bold> Participant was treated as a random intercept to control for inter-speaker differences in overall accuracy. In addition, by-participant random slopes for all of the word- and segment-level factors mentioned above were also included in the initial model. These by-participant random slopes were motivated by the fact that vocabulary size&#8212;which may vary across individuals&#8212;has been shown to associate with the effect of lexical factors like neighborhood size and average neighborhood frequency (<xref ref-type="bibr" rid="B193">Yap, Sibley, Balota, Ratcliff, &amp; Rueckl, 2015</xref>).</p>
</sec>
<sec>
<title>6.1.3. Procedure</title>
<p>We began with an initial, full model incorporating all of the fixed and random effects described above. This model was then simplified by a standard step-down model-selection procedure making use of the <monospace>anova()</monospace> function and likelihood ratio test provided by R. This procedure, described in greater detail in Appendix B, resulted in the final, best model in (3), where (1&#124;F) indicates a simple random effect of factor F.</p>
<table-wrap>
<table content-type="example">
<tbody>
<tr>
<td>(3)</td>
<td>Best model</td>
</tr>
<tr>
<td>&#160;</td>
<td>Accuracy &#126; Stimulus Similarity + Category Similarity (Mean) + Functional Load + Distributional Overlap + Word Frequency + (1 &#124; Unordered Stimulus Pair) + (1 &#124; Participant)</td>
</tr>
</tbody>
</table>
</table-wrap>
</sec>
</sec>
<sec>
<title>6.2. Statistical results</title>
<sec>
<title>6.2.1. Unimportant factors</title>
<p>A number of fixed and random effects were dropped during the model selection procedure. The fixed effects which fell out of the model were Category Similarity (<italic>SD</italic>), Segmental Frequency, all but one of the word-level factors (Wordhood, Neighborhood Density, Average Neighborhood Frequency, and Bigram Frequency), and Response Time. We suspect that Category Similarity (<italic>SD</italic>) emerged as insignificant because Category Similarity (mean) provides a better estimate of the distance between two phonemic categories: Category Similarity (mean) reflects the distance between category centroids&#8212;a property which clearly impacts the overall similarity between two categories&#8212;while Category Similarity (<italic>SD</italic>) reflects the variability in pairwise token comparisons across those categories&#8212;a property which could lead to either more or less overlap between categories depending on the shape of the variation. Like Bundgaard-Nielsen and Baker (<xref ref-type="bibr" rid="B29">2014</xref>) and Bundgaard-Nielsen et al. (<xref ref-type="bibr" rid="B30">2015</xref>), we did not find an effect of Segmental Frequency. Most word-level predictors were dropped from our final model, which we interpret as evidence that word-level factors had a limited effect on discrimination accuracy in our study. This perhaps reflects the fact that listeners could carry out the task (AX discrimination) without accessing lexical items of Kaqchikel (see Sections 5.3.2, 8 for more discussion). The insignificance of Response Time further suggests that there was no meaningful speed-accuracy trade-off in this study.</p>
<p>The dropped random intercepts were Stimulus Order, Onset vs. Coda, and List. All of the by-participant random slopes for word-level factors also fell out of the final model. Unlike e.g., Bundgaard-Nielsen et al. (<xref ref-type="bibr" rid="B30">2015</xref>), we found no evidence that the order of presentation of the two stimuli within a pair affected participant responses. The insignificance of List suggests the stimuli were randomized successfully across participants, such that the distribution of stimuli within and across lists did not serve as an accidental confound. The failure to retain by-participant random slopes for word-level factors in the final model may owe to several factors: Either vocabulary size was fairly homogenous across participants, or differences in vocabulary size do not have a material effect on the strength of the statistically significant predictors (segment-level Functional Load and Distributional Overlap, and Word Frequency).</p>
</sec>
<sec>
<title>6.2.2. Explanatory factors</title>
<p>The significant fixed factors in the best model are reported in Table <xref ref-type="table" rid="T2">2</xref>.</p>
<table-wrap id="T2">
<label>Table 2</label>
<caption>
<p>Regression statistics of the fixed effects in the best model predicting response accuracy (correct: 1, incorrect: 0) in log odds space.</p>
</caption>
<table>
<tr>
<th align="left"></th>
<th align="center"><italic>&#946;</italic></th>
<th align="center">SE(&#946;)</th>
<th align="center">&#124;<italic>z</italic>&#124;</th>
<th align="center" colspan="2"><italic>p</italic>-value</th>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">(Intercept)</td>
<td align="right">0.8042</td>
<td align="right">0.1621</td>
<td align="right">4.963</td>
<td align="right">&lt;.001</td>
<td align="left">***</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Stimulus Similarity</td>
<td align="right">&#8211;1.0720</td>
<td align="right">0.1151</td>
<td align="right">9.316</td>
<td align="right">&lt;.001</td>
<td align="left">*</td>
</tr>
<tr>
<td align="left">Category Similarity (mean)</td>
<td align="right">&#8211;0.3876</td>
<td align="right">0.1238</td>
<td align="right">3.131</td>
<td align="right">&lt;.005</td>
<td align="left">**</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Functional Load</td>
<td align="right">0.4653</td>
<td align="right">0.1649</td>
<td align="right">2.822</td>
<td align="right">&lt;.005</td>
<td align="left">**</td>
</tr>
<tr>
<td align="left">Distributional overlap</td>
<td align="right">&#8211;0.6320</td>
<td align="right">0.1607</td>
<td align="right">3.933</td>
<td align="right">&lt;.001</td>
<td align="left">***</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Word Frequency (abs. diff.)</td>
<td align="right">0.1848</td>
<td align="right">0.1068</td>
<td align="right">1.731</td>
<td align="right">.084</td>
<td align="left"><sup>.</sup></td>
</tr>
</table>
<table-wrap-foot>
<fn><p>&#8216;*&#8217; stands for <italic>p</italic> &lt; .05; &#8216;**&#8217; for <italic>p</italic> &lt; .01; &#8216;***&#8217; for <italic>p</italic> &lt; .001; &#8216;.&#8217; for <italic>p</italic> &lt; .10; <italic><sup>n.s.</sup></italic> for &#8216;not significant.&#8217;</p></fn>
</table-wrap-foot>
</table-wrap>
<p>Both acoustic similarity measures are highly significant, particularly Stimulus Similarity. These acoustic measures have negative coefficients, meaning that the greater the acoustic similarity between two stimuli and their associated phonemic categories, the harder it is to discriminate those stimuli. Second, Functional Load has a positive coefficient, meaning that the higher the pairwise functional load of two stops, the easier it is to discriminate syllables differentiated by those stops. Third, Distributional Overlap has a negative coefficient, meaning that the more phonotactic environments shared by two stops, the harder it is to discriminate them (this was an unexpected finding, which we discuss in detail below). Fourth, the only remaining word-level factor, Word Frequency, has a positive coefficient, meaning that the bigger the difference in token frequency between two syllables which are also words of Kaqchikel, the easier it is to discriminate them. However, unlike the other predictors in this final model, the effect of Word Frequency is only marginally significant (<italic>p</italic> = .084), consistent with our overall finding that word-level factors do not have much of an effect on discrimination accuracy in our study. We believe that this effect of word frequency, though marginal, reflects the general importance of this factor in psycholinguistic processing: Word frequency is consistently the strongest word-level factor in lexical retrieval tasks (such as lexical decision tasks) in a wide range of languages (<xref ref-type="bibr" rid="B58">Ferrand et al., 2010</xref>; <xref ref-type="bibr" rid="B102">Keuleers, Diependaele, &amp; Brysbaert, 2010</xref>; <xref ref-type="bibr" rid="B103">Keuleers et al., 2012</xref>; <xref ref-type="bibr" rid="B169">Sze, Rickard Liow, &amp; Yap, 2014</xref>).</p>
<p>While all remaining fixed effects are statistically important according to our model selection procedure, differences in the size of the coefficients suggest that these predictors differ in their relative strength. Stimulus Similarity was the most important predictor (&#124;<italic>&#946;</italic>&#124; = 1.0720), followed by Distributional Overlap (&#124;<italic>&#946;</italic>&#124; = 0.6320), Functional Load (&#124;<italic>&#946;</italic>&#124; = 0.4653), Category Similarity (mean) (&#124;<italic>&#946;</italic>&#124; = 0.3876), and Word Frequency (&#124;<italic>&#946;</italic>&#124; = 0.1848).<xref ref-type="fn" rid="n12">12</xref></p>
</sec>
</sec>
</sec>
<sec>
<title>7. Interim discussion</title>
<p>The statistical analysis in Section 6 established that both Stimulus Similarity and Category Similarity had an effect on discriminability in our perception study. The effect of Stimulus Similarity is unsurprising&#8212;stimuli that were acoustically more similar were, expectedly, harder to discriminate. The effect of Category Similarity requires additional interpretation.</p>
<p>Recall that Category Similarity was computed on the basis of acoustic similarity between stop categories as they occur in our corpus of spontaneous spoken Kaqchikel (Sections 4.1, 5.2). We believe that this corpus provides a good approximation of the acoustic properties of Kaqchikel stops as they occur in actual, fluent speech. As such, we take the significant effect of Category Similarity as an indication that phonemic categories which are acoustically well-separated in regular Kaqchikel speech are easier to discriminate in perceptual tasks.</p>
<p>At the theoretical level, this finding suggests that consonant discrimination is mediated by some representation of prior phonetic experience. In particular, these results are consistent with the view that speakers possess mental representations for phonemic categories which include rich phonetic detail, including (at least) some information about the acoustic properties which are typically associated with actual productions of each phoneme category in everyday speech. This claim accords with exemplar models of lexical representation, which assume that linguistic units (words, phonemes, etc.) are represented as clouds of episodic memories, which store phonetic representations of specific instances on which those units were encountered in speech (e.g., <xref ref-type="bibr" rid="B63">Gahl &amp; Yu, 2006</xref>; <xref ref-type="bibr" rid="B72">Goldinger, 1996</xref>, <xref ref-type="bibr" rid="B73">1998</xref>; <xref ref-type="bibr" rid="B95">K. Johnson, 2005</xref>; <xref ref-type="bibr" rid="B141">Pierrehumbert, 2001</xref> and references there). Our results are also consistent with the alternative view that phonemic categories are represented in a more abstract, parametric fashion, as vectors of values along specific dimensions (e.g., VOT, closure duration, etc.) which are specified separately for each phonemic category (see <xref ref-type="bibr" rid="B56">Ernestus, 2014</xref>; <xref ref-type="bibr" rid="B143">Pierrehumbert, 2016</xref>; <xref ref-type="bibr" rid="B163">Smits, Sereno, &amp; Jongman, 2006</xref> for discussion).</p>
<p>We also found that two predictors related to the lexical structure of Kaqchikel&#8212;Functional Load and Distributional Overlap&#8212;made a significant contribution to predicting consonant confusions in our study. Notably, both of these factors are segment-level factors: The word-level factors considered here had essentially no effect on stop consonant confusions. Following Hall (<xref ref-type="bibr" rid="B78">2012</xref>) and others, we take Functional Load and Distributional Overlap to be expressions of a <italic>gradient</italic> notion of segment-level phonemic contrast. In this sense, the relative predictability of two phonemes across contexts, and the precise number of words distinguished by those phonemes, provide a scalar characterization of how contrastive those phonemes are (i.e., how much lexical &#8216;work&#8217; is done by the contrast between those phonemes). Our results suggest that contrasts which have a higher functional importance in Kaqchikel are also easier to discriminate, as indicated by the positive correlation between accuracy and functional load. This finding is consistent with the view that language-specific phonemic contrasts &#8216;warp&#8217; the perceptual space in both categorical and gradient ways (e.g., <xref ref-type="bibr" rid="B19">Boomershine et al., 2008</xref>; <xref ref-type="bibr" rid="B81">Hall &amp; Hume, in preparation</xref>; <xref ref-type="bibr" rid="B83">Hall et al., 2014</xref>; <xref ref-type="bibr" rid="B84">Harnsberger, 2000</xref>, <xref ref-type="bibr" rid="B85">2001a</xref>, <xref ref-type="bibr" rid="B86">2001b</xref>; <xref ref-type="bibr" rid="B99">Kataoka &amp; Johnson, 2007</xref> and references there).</p>
<p>This interpretation of the results is nonetheless complicated by the finding that distributional overlap is <italic>negatively</italic> correlated with accuracy in our study. Such a correlation indicates that phonemes which occur in more shared environments&#8212;that is, phonemes which are less predictable from context, and therefore <italic>more</italic> contrastive&#8212;are harder to discriminate. This effect is contrary to our finding for functional load, which suggests that segments which distinguish many word forms (and which are therefore <italic>not</italic> predictable from context) are easier to discriminate; it is also contrary to previous findings by Hall et al. (<xref ref-type="bibr" rid="B83">2014</xref>) and Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) on the effect of distributional overlap on segment discrimination.</p>
<p>We are uncertain as to the source of this discrepancy. We first considered whether the negative correlation could be driven by the behavior of the alveolar ejective /t<sup>&#660;</sup>/ alone. This segment has very low type and token frequencies in our corpora and in Kaqchikel more generally (Section 4.1.3), meaning that it should have a low degree of distributional overlap with other phones. Ejective /t<sup>&#660;</sup>/ is nonetheless highly perceptible&#8212;most d&#8242; values for comparisons involving /t<sup>&#660;</sup>/ are above 2, compared to a grand average of about 1.7 for all pairwise comparisons&#8212;and so this segment alone might be driving the negative correlation between accuracy and distributional overlap. However, this is not the case: When we re-run our analyses with comparisons involving /t<sup>&#660;</sup>/ excluded, the effect size of Distributional Overlap weakens, but the negative sign does not change (<italic>&#946;</italic> = &#8211;0.247, <italic>p</italic> &lt; .05).</p>
<p>Alternatively, this divergent result may reflect the methods used to calculate distributional overlap in our study. As emphasized by Hall (<xref ref-type="bibr" rid="B78">2012</xref>), measures of distributional overlap and contextual predictability are highly sensitive to the definition of &#8216;context&#8217; used. For example, Kaqchikel has a process which devoices syllable-final /l/ to [l&#805;] (e.g., /<italic>loq&#8217;ob&#8217;&#228;l</italic>/loq<sup>&#660;</sup>o&#595;&#601;l/ &#8594; [loq<sup>&#660;</sup>o&#595;&#601;l&#805;] &#8216;blessing&#8217;). These two sounds are distributed completely predictably, but only if &#8216;context&#8217; can refer to right-hand environments and syllable structure [__X] (i.e., [l&#805;] is always followed by a syllable boundary, and [l] never is). If, instead, &#8216;context&#8217; refers only to the left-hand segment [X__], these sounds would appear to be at least partially contrastive and unpredictable (e.g., both can be preceded by [a], <italic>wach&#8217;alal</italic> [wat&#865;&#643;<sup>&#660;</sup>alal&#805;] &#8216;my family&#8217;).</p>
<p>Previous work on this topic has computed distributional overlap using highly-specific, pre-defined contexts which either (a) reflect the structure of experimental stimuli used in the study (e.g., [a__a] in <xref ref-type="bibr" rid="B83">Hall et al., 2014</xref>), or (b) reflect prior observations about the phonotactic contexts responsible for conditioning the distribution of sounds in the language under investigation (e.g., [__z]<sub>&#963;</sub> in French, <xref ref-type="bibr" rid="B81">Hall &amp; Hume, in preparation</xref>). Here, we used an inductive method (Jeffrey&#8217;s divergence) defined over all possible trigram windows to compute contextual predictability (Section 5.3.1). This methodological difference alone may have contributed substantially to the difference between our results and the results of Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>); Hall et al. (<xref ref-type="bibr" rid="B83">2014</xref>). With this in mind, we explored several other methods of computing contextual predictability, using bigram windows ([__X], [X__]) instead of trigram windows; using type rather than token frequencies (as in <xref ref-type="bibr" rid="B81">Hall &amp; Hume, in preparation</xref>; <xref ref-type="bibr" rid="B83">Hall et al., 2014</xref>); and including or excluding comparisons involving /t<sup>&#660;</sup>/. No combination of these methods yielded the expected positive correlation between distributional overlap and discriminability: In each case, the correlation was either non-significant or remained negative in sign.</p>
<p>These practical considerations aside, there is at least one other way to interpret the negative correlation between Distributional Overlap (degree of contrastiveness) and discriminability in our study, which again relies on exemplar dynamics. Two sounds which tend to occur in the same contexts may have more opportunities to be confused, particularly if word misperception is sensitive to statistical properties of the lexicon, including phonotactic well-formedness (see <xref ref-type="bibr" rid="B170">Tang, 2015</xref>). If &#8216;confusing&#8217; sound A for sound B means erroneously storing a token of A as a token of B in the exemplar space, then frequent confusions between A and B should have the effect of making the exemplar clouds for A and B more similar over time (e.g., <xref ref-type="bibr" rid="B181">Wedel, 2004</xref>; see also <xref ref-type="bibr" rid="B138">Ohala, 1993</xref>). In this way, increased distributional overlap between two sounds could indirectly lead to greater confusability between those sounds by increasing the amount of overlap between their associated exemplar clouds.<xref ref-type="fn" rid="n13">13</xref> Choosing between these possible interpretations of the effect of Distributional Overlap remains an open question for future research.</p>
<p>To reiterate, the finding that high contextual predictability (low degree of contrast) leads to greater discriminability conflicts with both our theoretical expectations (Section 5.3.1) and past results on this question (<xref ref-type="bibr" rid="B81">Hall &amp; Hume, in preparation</xref>; <xref ref-type="bibr" rid="B83">Hall et al., 2014</xref>). We are unsure how to interpret this result, though we note that the existence of a positive correlation between contrastiveness and discriminability has not yet been conclusively established: Such a result is reported by Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>); Hall et al. (<xref ref-type="bibr" rid="B83">2014</xref>), but Hall (<xref ref-type="bibr" rid="B77">2009</xref>) finds no meaningful correlation at all between contrastiveness and discriminability (though Hall also discusses some potential issues which may have led to this null result).</p>
<p>In any case it seems clear, particularly for functional load, that gradient contrast has an effect on speech perception. However, the precise mechanism(s) behind patterns of contrast-driven perceptual warping remain somewhat obscure (see again <xref ref-type="bibr" rid="B99">Kataoka &amp; Johnson, 2007</xref>). In the following section we test two hypotheses which attempt to provide an explicit link between consonant discrimination and segment-level distributional measures (Functional Load and Distributional Overlap) in Kaqchikel.</p>
<p>The first hypothesis is that functional load and distributional overlap are computed online in speech perception tasks such as ours, and that these computations can affect real-time speech processing. We do not think that this hypothesis is likely to be correct: Speech perception is rapid and automatic, while the computation of functional load and distributional overlap should require substantial processing time, even if computed over some subset of the lexicon (see e.g., <xref ref-type="bibr" rid="B108">Kingston et al., 2016</xref>; <xref ref-type="bibr" rid="B125">McClelland &amp; Elman, 1986</xref>; <xref ref-type="bibr" rid="B126">McClelland, Mirman, &amp; Holt, 2006</xref>; <xref ref-type="bibr" rid="B127">McClelland, Rumelhart, &amp; Hinton, 1986</xref>; <xref ref-type="bibr" rid="B135">Norris, McQueen, &amp; Cutler, 2000</xref> and references there for discussion). We nonetheless believe that this hypothesis is worthy of some consideration.</p>
<p>Our second hypothesis is that functional load and distributional overlap condition speech perception by shaping low-level perceptual tuning during development. By &#8216;perceptual tuning,&#8217; we refer to the fact that listeners selectively attend to those phonetic dimensions which are informative and reliable for the discrimination of phonemic categories in their native language (<xref ref-type="bibr" rid="B43">L. Davidson et al., 2007</xref>; <xref ref-type="bibr" rid="B92">Holt &amp; Lotto, 2006</xref>; <xref ref-type="bibr" rid="B129">McGuire, 2007</xref>).</p>
<p>In the following section we attempt to disentangle these two hypotheses by investigating the timecourse of segment-level distributional factors (Functional Load and Distributional Overlap) in our study.</p>
</sec>
<sec>
<title>8. The timecourse of experience-based effects</title>
<p>Speech perception can be decomposed into at least three distinct tasks: auditory/acoustic processing; phone-level processing; and lexical retrieval (e.g., <xref ref-type="bibr" rid="B3">Babel &amp; Johnson, 2010</xref>; <xref ref-type="bibr" rid="B60">Fox, 1984</xref>; <xref ref-type="bibr" rid="B146">Pisoni, 1975</xref>; <xref ref-type="bibr" rid="B147">Pisoni &amp; Tash, 1974</xref>; <xref ref-type="bibr" rid="B148">Pitt &amp; Samuel, 1993</xref>; <xref ref-type="bibr" rid="B184">Werker &amp; Logan, 1985</xref> and references there). The first of these tasks, auditory/acoustic processing, involves mechanisms which are basically physiological in nature. As such, this aspect of speech processing is not expected to be substantially affected by the listener&#8217;s native language. Phone-level processing (sometimes called &#8216;phonetic&#8217; or &#8216;phonemic&#8217; processing, e.g., <xref ref-type="bibr" rid="B145">Pisoni, 1973</xref>; <xref ref-type="bibr" rid="B184">Werker &amp; Logan, 1985</xref>; <xref ref-type="bibr" rid="B186">Werker &amp; Tees, 1984b</xref>) involves the categorization of speech sounds into appropriate phonemes and/or allophones. This type of processing differs from auditory/acoustic processing in that it is necessarily conditioned by the listener&#8217;s native language, and is therefore expected to show sensitivity to past linguistic experience. Such sensitivity is also expected for any aspect of speech processing that involves lexical access, as languages (and speakers) obviously differ in their vocabularies.</p>
<p>Researchers disagree as to the relative independence of each of these aspects of speech processing (see <xref ref-type="bibr" rid="B106">Kingston, 2005a</xref>; <xref ref-type="bibr" rid="B108">Kingston et al., 2016</xref>; <xref ref-type="bibr" rid="B125">McClelland &amp; Elman, 1986</xref>; <xref ref-type="bibr" rid="B126">McClelland et al., 2006</xref>, <xref ref-type="bibr" rid="B127">1986</xref>; <xref ref-type="bibr" rid="B135">Norris et al., 2000</xref> for discussion and further references). There is nonetheless a broad consensus that native-language influences on speech perception emerge relatively late in the timecourse of speech processing. This is particularly true of lexical effects, which tend to influence speech perception sometime after the initiation of phone-level processing (<xref ref-type="bibr" rid="B60">Fox, 1984</xref>; though cf. <xref ref-type="bibr" rid="B108">Kingston et al., 2016</xref> and work cited there).</p>
<p>Assuming that these stages of speech processing have the rough temporal sequencing suggested by prior work (acoustic/auditory &#8658; phone-level &#8658; lexical), we can at least tentatively diagnose the mechanism behind the segment-level statistical effects in our study (Functional Load and Distributional Overlap) by investigating when in the course of speech processing those effects arise. If Functional Load (FL) and Distributional Overlap (DO) are computed online during speech perception, through some process of lexical access or lexical sampling, then the effect of these predictors should emerge relatively late. We would then expect stronger effects of FL/DO at slower response times. If, on the other hand, FL/DO affect speech processing by shaping low-level perceptual tuning (e.g., cue weighting) during acquisition, then we might expect to see the influence of these predictors even at relatively fast response times.</p>
<sec>
<title>8.1. Predictions</title>
<p>Among the significant predictors in our final model (3), Stimulus Similarity corresponds most closely to the kind of information that would be processed during the acoustic/auditory stage of speech perception. We thus expect that Stimulus Similarity should have a robust effect on response accuracy at even the fastest response times. Category Similarity, a measure which refers to language-specific phonetic distributions associated with individual phoneme categories, should emerge no earlier than than the purely acoustic measure of Stimulus Similarity. If functional load and distributional overlap are computed online, they should begin to affect responses at a later stage than either Stimulus Similarity or Category Similarity.</p>
<p>Apart from the relative onset of these effects, we might also find differences in how the influence of each factor changes over time. Even if all of the factors in the model begin to affect response accuracy at about the same point, some factors might still grow in strength over time, while others weaken instead. In particular, factors involving lexical access might be more evident at longer response latencies, under the assumption that the strength of lexical activation gradually increases over time, such that lexical factors influence phone-level activation more strongly at later stages of processing (e.g., <xref ref-type="bibr" rid="B108">Kingston et al., 2016</xref>).</p>
</sec>
<sec>
<title>8.2. Statistical modeling</title>
<p>To carry out a timecourse analysis of our results we fit a new regression model based on our previous best model (3). Five interaction terms were added to test whether the significant fixed effects in (3) interact with response time in predicting participant accuracy. Response Time was also added as a fixed effect, consistent with the standard practice of including simple effects for any predictor included in an interaction term. Nested model comparison shows that model fit is significantly improved when these five interaction terms are included (Table <xref ref-type="table" rid="T3">3</xref>; &#967;<sup>2</sup>(5) = 23.6, <italic>p</italic> &lt; .001).</p>
<table-wrap id="T3">
<label>Table 3</label>
<caption>
<p>Regression statistics of the best model (3) with added interaction terms between the five fixed effects and by-participant response time predicting response accuracy (correct: 1, incorrect: 0) in log odds space.</p>
</caption>
<table>
<tr>
<th align="left"></th>
<th align="center"><italic>&#946;</italic></th>
<th align="center">SE(&#946;)</th>
<th align="center">&#124;<italic>z</italic>&#124;</th>
<th align="center" colspan="2"><italic>p</italic>-value</th>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">(Intercept)</td>
<td align="right">0.8248</td>
<td align="right">0.1635</td>
<td align="right">5.044</td>
<td align="right">&lt;.001</td>
<td align="left">*</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Stimulus Similarity</td>
<td align="right">&#8211;1.0705</td>
<td align="right">0.1165</td>
<td align="right">9.193</td>
<td align="right">&lt;.001</td>
<td align="left">***</td>
</tr>
<tr>
<td align="left">Category Similarity (mean)</td>
<td align="right">&#8211;0.3975</td>
<td align="right">0.1257</td>
<td align="right">3.163</td>
<td align="right">.002</td>
<td align="left">**</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Functional Load</td>
<td align="right">0.4820</td>
<td align="right">0.1668</td>
<td align="right">2.889</td>
<td align="right">.004</td>
<td align="left">**</td>
</tr>
<tr>
<td align="left">Distributional Overlap</td>
<td align="right">&#8211;0.6544</td>
<td align="right">0.1630</td>
<td align="right">4.015</td>
<td align="right">&lt;.001</td>
<td align="left">***</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Word Frequency (abs. diff.)</td>
<td align="right">0.1705</td>
<td align="right">0.1084</td>
<td align="right">1.573</td>
<td align="right">.116</td>
<td align="left"><sup><italic>n.s.</italic></sup></td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Response Time</td>
<td align="right">&#8211;0.1437</td>
<td align="right">0.0637</td>
<td align="right">2.254</td>
<td align="right">.024</td>
<td align="left">*</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Stimulus Similarity:Response Time</td>
<td align="right">0.1877</td>
<td align="right">0.0742</td>
<td align="right">2.528</td>
<td align="right">.011</td>
<td align="left">*</td>
</tr>
<tr>
<td align="left">Category Similarity (mean):Response Time</td>
<td align="right">0.1591</td>
<td align="right">0.0794</td>
<td align="right">2.005</td>
<td align="right">.045</td>
<td align="left">*</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Functional Load:Response Time</td>
<td align="right">&#8211;0.2737</td>
<td align="right">0.1063</td>
<td align="right">2.575</td>
<td align="right">.010</td>
<td align="left">*</td>
</tr>
<tr>
<td align="left">Distributional Overlap:Response Time</td>
<td align="right">0.3389</td>
<td align="right">0.1076</td>
<td align="right">3.149</td>
<td align="right">.002</td>
<td align="left">**</td>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left">Word Frequency (abs. diff.):Response Time</td>
<td align="right">0.0132</td>
<td align="right">0.0686</td>
<td align="right">0.193</td>
<td align="right">.8465</td>
<td align="left"><sup><italic>n.s.</italic></sup></td>
</tr>
</table>
</table-wrap>
<p>We also performed a separate timecourse analysis treating Response Time as a discrete variable rather than a continuous one. Dichotomizing continuous variables is often discouraged (<xref ref-type="bibr" rid="B2">Baayen, 2008, p. 259</xref>), but we performed this additional analysis because it more closely resembles the treatment of timecourse effects on speech processing in some previous work (<xref ref-type="bibr" rid="B3">Babel &amp; Johnson, 2010</xref>; <xref ref-type="bibr" rid="B106">Kingston, 2005a</xref>). First, each participant&#8217;s responses were divided into three equally-sized bins (i.e., by-participant terciles): fast responses (Early), medium-speed responses (Middle), and slow responses (Late). The mean response times for each bin (across participants) were about 400 ms, 650 ms, and 1200 ms; the first bin falls roughly in the range of response times associated with auditory processing, while the latter two fall in the range associated with phone-level and/or lexical processing (e.g., <xref ref-type="bibr" rid="B3">Babel &amp; Johnson, 2010</xref>; <xref ref-type="bibr" rid="B60">Fox, 1984</xref>; <xref ref-type="bibr" rid="B184">Werker &amp; Logan, 1985</xref>; <xref ref-type="bibr" rid="B186">Werker &amp; Tees, 1984b</xref>). We Helmert-coded this discrete, three-level timecourse predictor, and re-fit the model (3) using the same structure used for our continuous response time predictor. This analysis yielded the same qualitative results as the analysis which treated timecourse as a continuous predictor. We report only the continuous model below.<xref ref-type="fn" rid="n14">14</xref></p>
<p>The significant interactions reported in Table <xref ref-type="table" rid="T3">3</xref> suggest that the strength of our acoustic and distributional predictors did vary as a function of response time. To dig deeper into the interaction between these factors and response time, we fit a separate regression model for each of the response time tercile bins described above (Table <xref ref-type="table" rid="T4">4</xref>), using the same model structure (3) which we used in the analysis of the full data set.</p>
<table-wrap id="T4">
<label>Table 4</label>
<caption>
<p>Regression statistics for the fixed effects of three regression models, computed over by-participant response time terciles, predicting response accuracy (correct: 1, incorrect: 0) in log odds space.</p>
</caption>
<table>
<tr>
<th align="left" valign="top" rowspan="3"></th>
<th align="center" valign="top" colspan="2">Early<break/>(<italic>&#956;</italic> &#8776; 400 ms)</th>
<th align="center" valign="top" colspan="2">Middle<break/>(<italic>&#956;</italic> &#8776; 650 ms)</th>
<th align="center" valign="top" colspan="2">Late<break/>(<italic>&#956;</italic> &#8776; 1200 ms)</th>
</tr>
<tr>
<th colspan="6"><hr/></th>
</tr>
<tr>
<th align="center" colspan="2"><italic>&#946;</italic></th>
<th align="center" colspan="2"><italic>&#946;</italic></th>
<th align="center" colspan="2"><italic>&#946;</italic></th>
</tr>
<tr>
<td colspan="7"><hr/></td>
</tr>
<tr>
<td align="left">Stimulus Similarity</td>
<td align="right">&#8211;1.4515</td>
<td align="left">***</td>
<td align="right">&#8211;1.1651</td>
<td align="left">***</td>
<td align="right">&#8211;0.7465</td>
<td align="left">***</td>
</tr>
<tr>
<td align="left">Category Similarity</td>
<td align="right">&#8211;0.6544</td>
<td align="left">**</td>
<td align="right">&#8211;0.3020</td>
<td align="left">.</td>
<td align="right">&#8211;0.2876</td>
<td align="left">*</td>
</tr>
<tr>
<td colspan="7"><hr/></td>
</tr>
<tr>
<td align="left">Functional Load</td>
<td align="right">0.9001</td>
<td align="left">**</td>
<td align="right">0.4116</td>
<td align="left">.</td>
<td align="right">0.2853</td>
<td align="left">.</td>
</tr>
<tr>
<td align="left">Distributional Overlap</td>
<td align="right">&#8211;1.1437</td>
<td align="left">***</td>
<td align="right">&#8211;0.8765</td>
<td align="left">***</td>
<td align="right">&#8211;0.2797</td>
<td align="left">.</td>
</tr>
<tr>
<td colspan="7"><hr/></td>
</tr>
<tr>
<td align="left">Word Frequency (abs. diff.)</td>
<td align="right">0.2671</td>
<td align="left"><sup><italic>n.s.</italic></sup></td>
<td align="right">0.2314</td>
<td align="left"><sup><italic>n.s.</italic></sup></td>
<td align="right">0.0607</td>
<td align="left"><sup><italic>n.s.</italic></sup></td>
</tr>
</table>
</table-wrap>
</sec>
<sec>
<title>8.3. Statistical results</title>
<p>Table <xref ref-type="table" rid="T3">3</xref> presents the regression statistics of the model with interaction terms between the five fixed effects and Response Time. We first note that Word Frequency and its interaction term with Response Time do not reach statistical significance. The other four interaction terms do reach statistical significance. Crucially, the coefficients of these four significant interaction terms indicate that the four fixed effects decrease in strength as response times increase.</p>
<p>Table <xref ref-type="table" rid="T4">4</xref> presents the regression statistics for each of the three regression models, grouped by response time bin. We first note that Word Frequency does not reach statistical significance in any response time bin. The other four predictors&#8212;Stimulus Similarity, Category Similarity, Functional Load, and Distributional Overlap&#8212;have consistent effects across all three response time bins. Each of these four predictors influence response accuracy even in the earliest bin. Additionally, all of these predictors (including the insignificant predictor Word Frequency) decrease in strength as response times increase. This decrease in strength is evident in both the magnitude of the effects (the values of the <italic>&#946;</italic> coefficients) and the level of statistical significance reached. Together these findings suggest that all four significant factors kick in early, but decrease in strength over time.</p>
</sec>
<sec>
<title>8.4. Interpretation of timecourse analysis</title>
<p>This timecourse analysis shows that three experience-based factors (Category Similarity, Functional Load, and Distributional Overlap) began to affect discrimination at about the same early timepoint as the acoustic/auditory factor Stimulus Similarity.</p>
<p>For present purposes, the most important result is that Functional Load and Distributional Overlap influenced response accuracy even at very fast response times (those in the Early tercile bin). This result is consistent with the view that these two &#8216;lexical&#8217; measures impinge on speech perception somewhat indirectly, most likely by influencing which acoustic dimensions speakers attend to more closely to during phone-level processing. If Functional Load and Distributional Overlap condition speech perception through perceptual tuning, these factors are <italic>expected</italic> to show the same timecourse as Category Similarity, as all three measures reflect perceptual processes which in some way refer to the phonetic dimensions that distinguish phoneme categories in the listener&#8217;s native language.</p>
<p>We cannot completely rule out the possibility that Functional Load and Distributional Overlap are computed online during speech perception. For one, the inter-stimulus interval in this study (up to 800 ms) may have been sufficiently long that listeners were able to carry out some form of lexical access even for fairly quick responses. Nevertheless, we believe that this interpretation of our results is at odds with several observations. First, the AX discrimination task used in this study neither required nor encouraged lexical access, particularly because most stimuli were not actual words of Kaqchikel. Second, we think it is inherently unlikely that listeners carry out the large-scale lexical access that would be needed to accurately compute measures like functional load online. Speech processing is simply too rapid to involve lexical access at this scale during real-time listening. This argument is bolstered by the observation that the effect of these &#8216;lexical&#8217; measures emerged early and <italic>weakened</italic> over time; were some form of bulk lexical access involved, we should expect to see the strength of these measures increase over time instead, as more of the lexicon is accessed and analyzed.</p>
</sec>
</sec>
<sec>
<title>9. Discussion</title>
<sec>
<title>9.1. Theoretical contributions</title>
<p>The core theoretical contributions of this article are twofold. First, we have demonstrated experimentally that prior linguistic experience affects speech perception, not simply because different languages have different phonemic inventories (e.g., <xref ref-type="bibr" rid="B184">Werker &amp; Logan, 1985</xref>; <xref ref-type="bibr" rid="B185">Werker &amp; Tees, 1984a</xref>), but also because languages differ in the fine phonetic details associated with phonemic categories, as well as in their lexical structure and patterns of usage. These results replicate and extend past research showing that highly specific statistical patterns in a listener&#8217;s native language can have extensive effects on perceptual processing, even in experimental tasks that do not obviously require lexical access.</p>
<p>Importantly, our investigation has established these results in the context of a language&#8212;Kaqchikel Maya&#8212;which is sociolinguistically and structurally very different from the majority languages which are most often studied in speech perception research (Section 1). We hope that this work will encourage researchers to continue expanding the speech perception literature, and the phonetics literature more generally, to include a wider range of lesser-studied languages. Only in this way can we establish cross-linguistically valid theories of speech perception and production.</p>
</sec>
<sec sec-type="methods">
<title>9.2. Methodological contributions</title>
<p>In this study, we replicated several findings of experience-based effects in speech perception which have previously been demonstrated for majority languages using much richer resources (e.g., substantially larger written corpora, Section 4.2.2). We take this result to be an indirect validation of the use of small corpora in speech perception studies. Despite their shortcomings, small, noisy corpora can make valuable contributions to speech perception research, provided they are carefully processed beforehand. Our results also supply a positive answer to the general question of whether reliable speech perception research can be conducted in the field, outside of highly-controlled laboratory settings (<xref ref-type="bibr" rid="B187">Whalen &amp; McDonough, 2015</xref>; see also <xref ref-type="bibr" rid="B45">DiCanio, 2014</xref> for an excellent recent example of this kind of research).</p>
<p>To further support this claim, we now compare our findings to a similar study conducted with a majority language (French) in a laboratory setting. Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) investigated segment-level statistical effects on the discriminability of French vowels. They considered many of the same predictors we investigated in our study, including Stimulus Similarity, Functional Load, Distributional Overlap, and Segmental Token Frequency; this parallelism allows us to compare the two studies rather directly.</p>
<p>With the exception of Category Similarity and Word Frequency (which were not examined by Hall and Hume), the significant predictors in our study (Stimulus Similarity, Functional Load, and Distributional Overlap) were also statistically significant predictors of vowel discrimination in Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) (though the direction of the effect for Distributional Overlap was different in the two studies). Segment Token Frequency did not reach significance in either our study or in Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>). In terms of the relative importance of these predictors, both studies found that Stimulus Similarity played the strongest role, followed by predictors related to phonological contrastiveness (Functional Load and Distributional Overlap), with Distributional Overlap being a better predictor of response accuracy than Functional Load (though only when /t<sup>&#660;</sup>/ is included in the analysis). We find the parallelism between these results to be rather encouraging, especially given the following differences between the two studies:</p>
<list list-type="order">
<list-item><p><bold>Task:</bold> Hall and Hume used a multiple forced-choice identification task, while our study used an AX discrimination task.</p></list-item>
<list-item><p><bold>Target segments:</bold> Hall and Hume examined vowels, while our study examined consonants.</p></list-item>
<list-item><p><bold>Stimulus presentation:</bold> Hall and Hume presented their stimuli without any masking noise, while we presented our stimuli in speech-shaped noise at a 0 dB SNR.</p></list-item>
<list-item><p><bold>Experimental setting:</bold> Hall and Hume tested their participants in a controlled laboratory setting in a sound-attenuated booth, while we tested our participants in a quiet room which was not sound-attenuated.</p></list-item>
<list-item><p><bold>Language:</bold> Hall and Hume examined French, a Romance language, while we examined Kaqchikel, a Mayan language.</p></list-item>
<list-item><p><bold>Culture:</bold> the participants in Hall and Hume&#8217;s study were likely to have some experience with psychological experiments, and extensive experience with computers. Our participants did not in general have such experience.</p></list-item>
<list-item><p><bold>Quality of the corpus estimates:</bold> the segment-level predictors in Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) were estimated using a written corpus that is very large (65.1 million words), well-balanced, and highly speech-like (compiled from books, and subtitles of films and TV shows). Our segment-level predictors were estimated using a written corpus that is small (0.7 million words), unbalanced, and not particularly speech-like (mostly governmental and religious documents).</p></list-item>
</list>
<p>Despite these differences, which are varied and numerous, the two studies arrive at strikingly similar conclusions about the kinds of predictors which affect speech perception, as well as their relative importance. This comparison thus reaffirms our claim that conducting speech perception research in the field can result in findings that are comparable to those done in a laboratory.</p>
<p>However, it should also be noted that our results are substantially &#8216;noisier&#8217; than the results of Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>). In particular, the fixed effects components in our statistical model (Section 6) capture 23.2% of the variance in our data; in contrast, the fixed effects components in the statistical model reported in Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>) capture as much as 82% of the variance in their study.<xref ref-type="fn" rid="n15">15</xref> It is unclear to us which differences between the studies (including, but not limited to those outlined above) could have led to such a dramatic difference in model fit. Answering this question would require several follow-up studies which reduce the methodological differences with Hall and Hume (<xref ref-type="bibr" rid="B81">in preparation</xref>), a project we leave for future research.</p>
</sec>
</sec>
<sec sec-type="supplementary-material">
<title>Additional Files</title>
<p>The additional files for this article can be found as follows:</p>
<supplementary-material id="S1" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.100.s1">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-9-100-s1.pdf">labphon-9-100-s1.pdf</inline-supplementary-material>]-->
<label>Appendix A</label>
<caption>
<p>Processing of written corpus. DOI: <uri>https://doi.org/10.5334/labphon.100.s1</uri></p>
</caption>
</supplementary-material>
<supplementary-material id="S2" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.100.s2">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-9-100-s2.pdf">labphon-9-100-s2.pdf</inline-supplementary-material>]-->
<label>Appendix B</label>
<caption>
<p>Model construction and selection. DOI: <uri>https://doi.org/10.5334/labphon.100.s2</uri></p>
</caption>
</supplementary-material>
<supplementary-material id="S3" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.100.s3">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-9-100-s3.pdf">labphon-9-100-s3.pdf</inline-supplementary-material>]-->
<label>Appendix C</label>
<caption>
<p>Non-significant predictors in the AX discrimination study. DOI: <uri>https://doi.org/10.5334/labphon.100.s3</uri></p>
</caption>
</supplementary-material>
<supplementary-material id="S4" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.100.s4">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-9-100-s4.pdf">labphon-9-100-s4.pdf</inline-supplementary-material>]-->
<label>Appendix D</label>
<caption>
<p>d&#8242; scores for all target stop contrasts. DOI: <uri>https://doi.org/10.5334/labphon.100.s4</uri></p>
</caption>
</supplementary-material>
</sec>
</body>
<back>
<fn-group>
<fn id="n1"><p>A few studies have examined functional load in languages with fairly agglutinative morphology, such as Japanese, Korean, and Swahili (<xref ref-type="bibr" rid="B136">Oh, Coup&#233;, Marsico, &amp; Pellegrino, 2015</xref>; <xref ref-type="bibr" rid="B137">Oh, Pellegrino, Coup&#233;, &amp; Marsico, 2013</xref>). To our knowledge, no existing studies have addressed the relationship between functional load and speech perception in languages of this morphological type.</p></fn>
<fn id="n2"><p>The speech-shaped noise was generated from a four-hour acoustic corpus of spontaneous spoken Kaqchikel (Section 4.1), using a Praat script in the library praat-semiauto (<xref ref-type="bibr" rid="B128">McCloy, 2014</xref>). This Praat script took a directory of .wav files extracted from the spoken corpus and generated a Gaussian noise file which was spectrally shaped to match the long-term average spectrum of that corpus (essentially following <xref ref-type="bibr" rid="B150">Quen&#233; &amp; van Delft, 2010</xref>).</p></fn>
<fn id="n3"><p>d&#8242; scores are <italic>z</italic>-scores, so a d&#8242; of 1 would be obtained for a participant who responded with roughly 69% accuracy on both same and different trials, and a d&#8242; of 1.5 would be obtained for a participant who responded with a bit more than 77% accuracy on both same and different trials.</p></fn>
<fn id="n4"><p>In addition to computing d&#8242; scores over stop combinations, we computed a d&#8242; score for each participant. One participant had a very low d&#8242; score (0.047), more than 3 standard deviations from the mean d&#8242; score across participants. We re-ran our best statistical model (3) with this participant&#8217;s results excluded, and the model statistics remained virtually the same.</p></fn>
<fn id="n5"><p>This suite of Python scripts is currently not available, as they are being developed as part of another ongoing project.</p></fn>
<fn id="n6"><p>As a rough assessment of the accuracy of our forced alignment model, we hand-corrected a subset of the TextGrids produced by forced alignment, and compared them to the original, automatically aligned output. On average, stop consonants were well-identified by our alignment model: Out of 428 stops, the median alignment error was 10 ms (mean = 16 ms) (see also <xref ref-type="bibr" rid="B46">DiCanio et al., 2013</xref>). Further, 25% of alignments agreed to the exact millisecond, and 86% of alignment errors were 20 ms or shorter in size. These errors appear to be more-or-less evenly distributed across stop types: Consequently, alignment errors are unlikely to have skewed our measures of acoustic similarity in any particular direction. We thank Andrea Maynard for carefully hand-correcting these TextGrids.</p></fn>
<fn id="n7"><p>Same trials are often taken into account in statistical analyses of discriminability based on d&#8242; (<xref ref-type="bibr" rid="B118">Macmillan &amp; Creelman, 2005</xref>), as a way of controlling for individual response biases. In our model, response biases are captured by including Participant as a random effect in the linear regression.</p>
<p>A 9-by-9 plot summarizing the d&#8242; scores for all target stop pairs, collapsed across vowel context and syllable position, is provided as an appendix (Appendix D).</p></fn>
<fn id="n8"><p>Our measure of category similarity (means and standard deviations of pooled DTW measurements) does not distinguish between [CV] and [VC] contexts. We initially considered computing category similarity separately for [CV] and [VC] contexts, to more closely match the experimental design of our perception study (Section 3). We abandoned this approach because of the sparseness of the acoustic corpus (e.g., the rarest phoneme /t<sup>&#660;</sup>/ only precedes or follows /&#39;e/ and /&#39;o/, and no other vowels). Coping with this problem would have required us to pool over other contextual properties, such as the quality of the adjacent vowel, and in doing so we would have ignored other perceptually relevant factors in the analysis.</p></fn>
<fn id="n9"><p>This measure of gradient contrastiveness is, confusingly, sometimes also known as &#8216;functional load&#8217; (e.g., <xref ref-type="bibr" rid="B104">King, 1967</xref>).</p></fn>
<fn id="n10"><p>JD is a symmetric variant of Kullback-Leibner distance (<xref ref-type="bibr" rid="B110">Kullback &amp; Leibler, 1951</xref>), which has also been used to measure distributional differences across phones, for instance in research on statistical learning of allophonic alternations (<xref ref-type="bibr" rid="B31">Calamaro &amp; Jarosz, 2015</xref>; <xref ref-type="bibr" rid="B140">Peperkamp, Le Calvez, Nadal, &amp; Dupoux, 2006</xref>).</p></fn>
<fn id="n11"><p>We have not yet performed a stability simulation (Section 4.2.2) assessing how reliably distributional overlap is estimated from corpora of different sizes. We nonetheless expect that distributional overlap can be reliably estimated from a fairly small corpus, like our corpus of Kaqchikel. Tang et al. (<xref ref-type="bibr" rid="B171">2015</xref>) found that estimates of functional load computed over syllable types reached stability even faster than estimates computed over word types: This likely reflects the fact that syllable types, being smaller units, are better represented in the corpus than word types. Since distributional overlap is computed over trigrams, which are similar in size to syllables, we also expect distributional overlap to be reliably estimated from a relatively small corpus.</p></fn>
<fn id="n12"><p>Wedel, Jackson, and Kaplan (<xref ref-type="bibr" rid="B182">2013</xref>); Wedel, Kaplan, and Jackson (<xref ref-type="bibr" rid="B183">2013</xref>) found that raw minimal pair counts (as a measure of functional load) were a better predictor of diachronic patterns of phoneme merger than lexical &#916;-entropy, the metric employed here. Our implementation of functional load (systemic &#916;-entropy) is proportional, but not identical, to the number of minimal pairs distinguished by a phoneme pair.</p>
<p>There is no current consensus as to which functional load metric is best, or whether a single metric is appropriate for analyzing all kinds of data. Given this uncertainty, we refitted our best model (3) using raw minimal pair counts and frequency-weighted minimal pair counts (as in <xref ref-type="bibr" rid="B182">Wedel, Jackson, &amp; Kaplan, 2013</xref>) instead of lexical &#916;-entropy. Functional load was not a significant predictor of consonant confusions under either of these alternative formulations (raw minimal pair counts: <italic>z</italic> = 0.268, <italic>p</italic> &gt; 0.788; frequency-weighted minimal pair counts: <italic>z</italic> = 0.323, <italic>p</italic> &gt; 0.746). To evaluate these two models against the model using lexical &#916;-entropy, we applied two model comparison metrics, AIC and BIC: These indicated that lexical &#916;-entropy provides the best fit for our data (&#916;-entropy: <sc>AIC</sc> = 2406, <sc>BIC</sc> = 2452; raw minimal pair counts: <sc>AIC</sc> = 2414, <sc>BIC</sc> = 2460; frequency-weighted minimal pair counts: <sc>AIC</sc> = 2414, <sc>BIC</sc> = 2460).</p></fn>
<fn id="n13"><p>A question that arises is why this potential effect of similarity in the exemplar space isn&#8217;t already captured by our measure of <sc>CATEGORY SIMILARITY</sc>, under the assumption that similarity in an acoustic corpus is a good reflection of similarity between abstract exemplar clouds (e.g., <xref ref-type="bibr" rid="B141">Pierrehumbert, 2001</xref>). One possibility is that distributional overlap affects the shape of exemplar distributions in a manner which is not captured by our relatively simple measure of category similarity (distance between centroids). Investigating this issue in detail would take us too far beyond the goals of the present article.</p></fn>
<fn id="n14"><p>The model treating response time as a continuous variable has a significantly better fit to our data than the model which uses discrete response time bins (continuous response time model: <sc>AIC</sc> = 2391, <sc>BIC</sc> = 2466; discrete response time model: <sc>AIC</sc> = 2399, <sc>BIC</sc> = 2479).</p></fn>
<fn id="n15"><p>The proportion of variance captured by fixed effects in our model was computed with the function <monospace>r.squaredGLMM()</monospace>, part of the MuMIn library in R (<xref ref-type="bibr" rid="B6">Barto&#324;, 2014</xref>; <xref ref-type="bibr" rid="B96">P. C. Johnson, 2014</xref>; <xref ref-type="bibr" rid="B133">Nakagawa &amp; Schielzeth, 2013</xref>). This function returns both marginal <inline-formula>
<alternatives>
<mml:math id="Eq005-mml"><mml:mrow><mml:msubsup><mml:mi>R</mml:mi><mml:mrow><mml:mtext mathvariant="italic">GLMM</mml:mtext></mml:mrow><mml:mn>2</mml:mn></mml:msubsup></mml:mrow></mml:math>
<tex-math id="M5">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
R_{GLMM}^2
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75177/"/>
</alternatives>
</inline-formula> and conditional <inline-formula>
<alternatives>
<mml:math id="Eq006-mml"><mml:mrow><mml:msubsup><mml:mi>R</mml:mi><mml:mrow><mml:mtext mathvariant="italic">GLMM</mml:mtext></mml:mrow><mml:mn>2</mml:mn></mml:msubsup></mml:mrow></mml:math>
<tex-math id="M6">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
R_{GLMM}^2
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75177/"/>
</alternatives>
</inline-formula>. The reported variance is the marginal <inline-formula>
<alternatives>
<mml:math id="Eq007-mml"><mml:mrow><mml:msubsup><mml:mi>R</mml:mi><mml:mrow><mml:mtext mathvariant="italic">GLMM</mml:mtext></mml:mrow><mml:mn>2</mml:mn></mml:msubsup></mml:mrow></mml:math>
<tex-math id="M7">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
R_{GLMM}^2
\]
\end{document}
</tex-math>
<graphic xlink:href="/article/id/6226/file/75177/"/>
</alternatives>
</inline-formula>, which represents the variance explained by fixed factors.</p></fn></fn-group>
<sec>
<title>Abbreviations</title>
<p>1 = first person; <sc>SG</sc> = singular; <sc>PL</sc> = plural; <sc>ASP</sc> = aspect; <sc>ABS</sc> = absolutive; <sc>ERG</sc> = ergative; <sc>DIR</sc> = directional; <sc>INCH</sc> = inchoative; <sc>CAUS</sc> = causative; <sc>TRANS</sc> = transitive; <sc>PASS</sc> = passive; <sc>NOM</sc> = nominalizer; <sc>SD</sc> = standard deviation.</p>
</sec>
<ack>
<title>Acknowledgements</title>
<p>First and foremost, we thank the Kaqchikel speakers who participated in the perception study described here, as well as in the development of our spoken and written corpora. Asociaci&#243;n Ceiba (<ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://www.ceibaguate.org.gt/">http://www.ceibaguate.org.gt/</ext-link>) kindly provided recording space for the production of our spoken corpus, which was collected with the assistance of the Comunidad Ling&#252;&#237;stica Kaqchikel (<ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://kaqchikel.almg.org.gt/">http://kaqchikel.almg.org.gt/</ext-link>). Centro Educativo Maya Aj Sya (<ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://mayaajsya.wordpress.com/about/">https://mayaajsya.wordpress.com/about/</ext-link>) gave us both space and extensive support for carrying out our perception experiment. <italic>Janila maty&#246;x chiwe iwonojel</italic>! We also thank Robert Henderson for logistical help in conducting the perception experiment. We are grateful to Doug Whalen (Haskins Laboratories), Jason Shaw (Yale University), and Uriel Cohen Priva (Brown University) for detailed feedback at various stages of the development of this project. In addition, we thank audiences at the Yale Phonetics/Phonology Reading Group, Speech Science Forum at University College London, Haskins Laboratories, the University of Hong Kong, Brown University, The Hong Kong Polytechnic University, The Education University of Hong Kong, UC Santa Cruz, <italic>Workshop on Structure and Constituency in Languages of the Americas 21, Sound Systems of Mexico and Central America II, Form and Analysis in Mayan Linguistics IV</italic>, the <italic>91st Annual Meeting of the Linguistic Society of America</italic>, and the <italic>24th Manchester Phonology Meeting</italic>.</p>
</ack>
<sec>
<title>Competing interests</title>
<p>The authors have no competing interests to declare.</p>
</sec>
<ref-list>
<ref id="B1"><label>1</label><mixed-citation publication-type="journal"><string-name><surname>Atkins</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Clear</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Ostler</surname>, <given-names>N.</given-names></string-name> <year>1992</year>. <article-title>Corpus design criteria</article-title>. <source>Literary and Linguistic Computing</source>, <volume>7</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>16</lpage>. DOI: <pub-id pub-id-type="doi">10.1093/llc/7.1.1</pub-id></mixed-citation></ref>
<ref id="B2"><label>2</label><mixed-citation publication-type="book"><string-name><surname>Baayen</surname>, <given-names>R.</given-names></string-name> <year>2008</year>. <source>Analyzing linguistic data: A practical introduction to statistics using R</source>. <publisher-loc>Cambridge, UK</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1017/CBO9780511801686</pub-id></mixed-citation></ref>
<ref id="B3"><label>3</label><mixed-citation publication-type="journal"><string-name><surname>Babel</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> <year>2010</year>. <article-title>Accessing psycho-acoustic perception and language-specific perception with speech sounds</article-title>. <source>Laboratory phonology</source>, <volume>1</volume>(<issue>1</issue>), <fpage>179</fpage>&#8211;<lpage>205</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/labphon.2010.009</pub-id></mixed-citation></ref>
<ref id="B4"><label>4</label><mixed-citation publication-type="journal"><string-name><surname>Baese-Berk</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Goldrick</surname>, <given-names>M.</given-names></string-name> <year>2009</year>. <article-title>Mechanisms of interaction in speech production</article-title>. <source>Language and cognitive processes</source>, <volume>24</volume>(<issue>4</issue>), <fpage>527</fpage>&#8211;<lpage>554</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/01690960802299378</pub-id></mixed-citation></ref>
<ref id="B5"><label>5</label><mixed-citation publication-type="thesis"><string-name><surname>Barrett</surname>, <given-names>R.</given-names></string-name> <year>1999</year>. <source>A grammar of Sipakapense Maya</source> (Unpublished doctoral dissertation). <publisher-name>University of Texas at Austin</publisher-name>.</mixed-citation></ref>
<ref id="B6"><label>6</label><mixed-citation publication-type="webpage"><string-name><surname>Barto&#324;</surname>, <given-names>K.</given-names></string-name> <year>2014</year>. <article-title>Mumin: Multi-model inference [Computer software manual]</article-title>. Retrieved from: <uri>http://CRAN.R-project.org/package=MuMIn</uri> (R package version 1.10.0).</mixed-citation></ref>
<ref id="B7"><label>7</label><mixed-citation publication-type="webpage"><string-name><surname>Bates</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Maechler</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Bolker</surname>, <given-names>B.</given-names></string-name> <year>2011</year>. <article-title>lme4: Linear mixed-effects models using S4 classes [Computer software manual]</article-title>. R package. (Version 0.999375-41, retrieved from: <uri>http://CRAN.R-project.org/package=lme4</uri>).</mixed-citation></ref>
<ref id="B8"><label>8</label><mixed-citation publication-type="journal"><string-name><surname>Benk&#237;</surname>, <given-names>J. R.</given-names></string-name> <year>2003</year>. <article-title>Analysis of English nonsense syllable recognition in noise</article-title>. <source>Phonetica</source>, <volume>60</volume>(<issue>2</issue>), <fpage>129</fpage>&#8211;<lpage>157</lpage>. DOI: <pub-id pub-id-type="doi">10.1159/000071450</pub-id></mixed-citation></ref>
<ref id="B9"><label>9</label><mixed-citation publication-type="webpage"><string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name> <year>2010</year>. <chapter-title>Contrast and laryngeal states in Tz&#8217;utujil</chapter-title>. In: <string-name><surname>McGuire</surname>, <given-names>G.</given-names></string-name> (ed.), <source>UC Santa Cruz Linguistics Research Center annual report</source>, <fpage>93</fpage>&#8211;<lpage>120</lpage>. <publisher-loc>Santa Cruz, CA</publisher-loc>: <publisher-name>LRC Publications</publisher-name>. (Available online at: <uri>http://people.ucsc.edu/gmcguir1/LabReport/BennettLRC.pdf</uri>).</mixed-citation></ref>
<ref id="B10"><label>10</label><mixed-citation publication-type="journal"><string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name> <year>2016</year>. <article-title>Mayan phonology</article-title>. <source>Language and Linguistics Compass</source>, <volume>10</volume>(<issue>10</issue>), <fpage>469</fpage>&#8211;<lpage>514</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/lnc3.12148</pub-id></mixed-citation></ref>
<ref id="B11"><label>11</label><mixed-citation publication-type="journal"><string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Ajsivinac Sian</surname>, <given-names>J.</given-names></string-name> (in preparation). <source>Un corpus fon&#233;tico del kaqchikel de Solol&#225;, Guatemala: narrativas espont&#225;neas</source>. Electronic corpus, recorded 2013.</mixed-citation></ref>
<ref id="B12"><label>12</label><mixed-citation publication-type="journal"><string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Coon</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Henderson</surname>, <given-names>R.</given-names></string-name> <year>2016</year>. <article-title>Introduction to Mayan linguistics</article-title>. <source>Language and Linguistics Compass</source>, <volume>10</volume>(<issue>10</issue>), <fpage>1</fpage>&#8211;<lpage>14</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/lnc3.12159</pub-id></mixed-citation></ref>
<ref id="B13"><label>13</label><mixed-citation publication-type="journal"><string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Tang</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Ajsivinac Sian</surname>, <given-names>J.</given-names></string-name> (in preparation). <article-title>Laryngeal co-occurrence restrictions as constraints on sub-segmental articulatory structure</article-title>.</mixed-citation></ref>
<ref id="B14"><label>14</label><mixed-citation publication-type="book"><string-name><surname>Best</surname>, <given-names>C. T.</given-names></string-name> <year>1995</year>. <chapter-title>A direct realist view of cross-language speech perception</chapter-title>. In: <string-name><surname>Strange</surname>, <given-names>W.</given-names></string-name> (ed.), <source>Speech perception and linguistic experience: Issues in cross-language research</source>, <fpage>171</fpage>&#8211;<lpage>204</lpage>. <publisher-loc>Timonium, MD</publisher-loc>: <publisher-name>York Press</publisher-name>.</mixed-citation></ref>
<ref id="B15"><label>15</label><mixed-citation publication-type="journal"><string-name><surname>Best</surname>, <given-names>C. T.</given-names></string-name>, <string-name><surname>McRoberts</surname>, <given-names>G.</given-names></string-name>, &amp; <string-name><surname>Goodell</surname>, <given-names>E.</given-names></string-name> <year>2001</year>. <article-title>Discrimination of non-native consonant contrasts varying in perceptual assimilation to the listener&#8217;s native phonological system</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>109</volume>(<issue>2</issue>), <fpage>775</fpage>&#8211;<lpage>794</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1332378</pub-id></mixed-citation></ref>
<ref id="B16"><label>16</label><mixed-citation publication-type="journal"><string-name><surname>Biber</surname>, <given-names>D.</given-names></string-name> <year>1993</year>. <article-title>Representativeness in corpus design</article-title>. <source>Literary and Linguistic Computing</source>, <volume>8</volume>(<issue>4</issue>), <fpage>243</fpage>&#8211;<lpage>257</lpage>. DOI: <pub-id pub-id-type="doi">10.1093/llc/8.4.243</pub-id></mixed-citation></ref>
<ref id="B17"><label>17</label><mixed-citation publication-type="book"><string-name><surname>Bladon</surname>, <given-names>A.</given-names></string-name> <year>1986</year>. <chapter-title>Phonetics for hearers</chapter-title>. In: <string-name><surname>McGregor</surname>, <given-names>G.</given-names></string-name> (ed.), <source>Language for hearers</source>, <fpage>1</fpage>&#8211;<lpage>24</lpage>. <publisher-loc>Oxford, UK</publisher-loc>: <publisher-name>Pergamon Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1093/llc/8.4.243</pub-id></mixed-citation></ref>
<ref id="B18"><label>18</label><mixed-citation publication-type="webpage"><string-name><surname>Boersma</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Weenink</surname>, <given-names>D.</given-names></string-name> <year>2016</year>. <source>Praat: Doing phonetics by computer (Version 6.0.23)</source>. Computer program. (Retrieved from: <uri>http://www.praat.org/</uri>).</mixed-citation></ref>
<ref id="B19"><label>19</label><mixed-citation publication-type="book"><string-name><surname>Boomershine</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name>, <string-name><surname>Hume</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> <year>2008</year>. <chapter-title>The impact of allophony versus contrast on speech perception</chapter-title>. In: <string-name><surname>Avery</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Dresher</surname>, <given-names>B.</given-names></string-name>, &amp; <string-name><surname>Rice</surname>, <given-names>K.</given-names></string-name> (eds.), <source>Contrast in phonology: Theory, perception, acquisition</source>, <fpage>145</fpage>&#8211;<lpage>171</lpage>. <publisher-loc>Berlin, Germany</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>.</mixed-citation></ref>
<ref id="B20"><label>20</label><mixed-citation publication-type="journal"><string-name><surname>Broadbent</surname>, <given-names>D.</given-names></string-name> <year>1967</year>. <article-title>Word-frequency effect and response bias</article-title>. <source>Psychological Review</source>, <volume>74</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>15</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/h0024206</pub-id></mixed-citation></ref>
<ref id="B21"><label>21</label><mixed-citation publication-type="thesis"><string-name><surname>Brody</surname>, <given-names>M.</given-names></string-name> <year>2004</year>. <source>The fixed word, the moving tongue: Variation in written Yucatec Maya and the meandering evolution toward unified norms</source> (Unpublished doctoral dissertation). <publisher-name>University of Texas Austin</publisher-name>, <publisher-loc>Austin, TX</publisher-loc>.</mixed-citation></ref>
<ref id="B22"><label>22</label><mixed-citation publication-type="journal"><string-name><surname>Browman</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Goldstein</surname>, <given-names>L.</given-names></string-name> <year>1986</year>. <article-title>Towards an articulatory phonology</article-title>. <source>Phonology yearbook</source>, <volume>3</volume>(<issue>21</issue>), <fpage>219</fpage>&#8211;<lpage>252</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0952675700000658</pub-id></mixed-citation></ref>
<ref id="B23"><label>23</label><mixed-citation publication-type="journal"><string-name><surname>Browman</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Goldstein</surname>, <given-names>L.</given-names></string-name> <year>1989</year>. <article-title>Articulatory gestures as phonological units</article-title>. <source>Phonology</source>, <volume>6</volume>(<issue>2</issue>), <fpage>201</fpage>&#8211;<lpage>251</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0952675700001019</pub-id></mixed-citation></ref>
<ref id="B24"><label>24</label><mixed-citation publication-type="journal"><string-name><surname>Browman</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Goldstein</surname>, <given-names>L.</given-names></string-name> <year>1992</year>. <article-title>Articulatory phonology: An overview</article-title>. <source>Phonetica</source>, <volume>49</volume>(<issue>3&#8211;4</issue>), <fpage>155</fpage>&#8211;<lpage>180</lpage>. DOI: <pub-id pub-id-type="doi">10.1159/000261913</pub-id></mixed-citation></ref>
<ref id="B25"><label>25</label><mixed-citation publication-type="journal"><string-name><surname>Brown</surname>, <given-names>C. R.</given-names></string-name>, &amp; <string-name><surname>Rubenstein</surname>, <given-names>H.</given-names></string-name> <year>1961</year>. <article-title>Test of response bias explanation of word-frequency effect</article-title>. <source>Science</source>, <volume>133</volume>(<issue>3448</issue>), <fpage>280</fpage>&#8211;<lpage>281</lpage>. DOI: <pub-id pub-id-type="doi">10.1126/science.133.3448.280</pub-id></mixed-citation></ref>
<ref id="B26"><label>26</label><mixed-citation publication-type="book"><string-name><surname>Brown</surname>, <given-names>R. M.</given-names></string-name>, <string-name><surname>Maxwell</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Little</surname>, <given-names>W.</given-names></string-name> <year>2010</year>. <source>La &#252;tz aw&#228;ch?: Introduction to Kaqchikel Maya language</source>. <publisher-loc>Austin, TX</publisher-loc>: <publisher-name>University of Texas Press</publisher-name>.</mixed-citation></ref>
<ref id="B27"><label>27</label><mixed-citation publication-type="journal"><string-name><surname>Brysbaert</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Diependaele</surname>, <given-names>K.</given-names></string-name> <year>2013</year>. <article-title>Dealing with zero word frequencies: A review of the existing rules of thumb and a suggestion for an evidence-based choice</article-title>. <source>Behavior Research Methods</source>, <volume>45</volume>(<issue>2</issue>), <fpage>422</fpage>&#8211;<lpage>430</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/s13428-012-0270-5</pub-id></mixed-citation></ref>
<ref id="B28"><label>28</label><mixed-citation publication-type="journal"><string-name><surname>Brysbaert</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>New</surname>, <given-names>B.</given-names></string-name> <year>2009</year>. <article-title>Moving beyond Ku&#269;era and Francis: A critical evaluation of current word frequency norms and the introduction of a new and improved word frequency measure for American English</article-title>. <source>Behavior research methods</source>, <volume>41</volume>(<issue>4</issue>), <fpage>977</fpage>&#8211;<lpage>990</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BRM.41.4.977</pub-id></mixed-citation></ref>
<ref id="B29"><label>29</label><mixed-citation publication-type="confproc"><string-name><surname>Bundgaard-Nielsen</surname>, <given-names>R. L.</given-names></string-name>, &amp; <string-name><surname>Baker</surname>, <given-names>B. J.</given-names></string-name> <year>2014</year>. <article-title>Frequency in the input affects perception of phonological contrasts for native speakers</article-title>. In: <string-name><surname>Hay</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Parnell</surname>, <given-names>E.</given-names></string-name> (eds.), <conf-name>Proceedings of the 15th Australasian International Speech Science and Technology Conference</conf-name>, <fpage>205</fpage>&#8211;<lpage>208</lpage>. <conf-loc>Christchurch, New Zealand</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association</conf-sponsor>.</mixed-citation></ref>
<ref id="B30"><label>30</label><mixed-citation publication-type="journal"><string-name><surname>Bundgaard-Nielsen</surname>, <given-names>R. L.</given-names></string-name>, <string-name><surname>Baker</surname>, <given-names>B. J.</given-names></string-name>, <string-name><surname>Kroos</surname>, <given-names>C. H.</given-names></string-name>, <string-name><surname>Harvey</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Best</surname>, <given-names>C. T.</given-names></string-name> <year>2015</year>. <article-title>Discrimination of multiple coronal stop contrasts in Wubuy (Australia): A natural referent consonant account</article-title>. <source>PLOS ONE</source>, <volume>10</volume>(<issue>12</issue>), <fpage>1</fpage>&#8211;<lpage>30</lpage>. DOI: <pub-id pub-id-type="doi">10.1371/journal.pone.0142054</pub-id></mixed-citation></ref>
<ref id="B31"><label>31</label><mixed-citation publication-type="journal"><string-name><surname>Calamaro</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Jarosz</surname>, <given-names>G.</given-names></string-name> <year>2015</year>. <article-title>Learning general phonological rules from distributional information: A computational model</article-title>. <source>Cognitive science</source>, <volume>39</volume>(<issue>3</issue>), <fpage>647</fpage>&#8211;<lpage>666</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/cogs.12167</pub-id></mixed-citation></ref>
<ref id="B32"><label>32</label><mixed-citation publication-type="book"><string-name><surname>Campbell</surname>, <given-names>L.</given-names></string-name> <year>1977</year>. <source>Quichean linguistic prehistory</source>, <fpage>81</fpage>. <publisher-loc>Berkeley, CA</publisher-loc>: <publisher-name>University of California Press</publisher-name>.</mixed-citation></ref>
<ref id="B33"><label>33</label><mixed-citation publication-type="book"><string-name><surname>Chacach Cutzal</surname>, <given-names>M.</given-names></string-name> <year>1990</year>. <chapter-title>Una descripci&#243;n fonol&#243;gica y morfol&#243;gica del kaqchikel</chapter-title>. In: <string-name><surname>England</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Elliott</surname>, <given-names>S.</given-names></string-name> (eds.), <source>Lecturas sobre la ling&#252;&#237;stica maya</source>, <fpage>145</fpage>&#8211;<lpage>190</lpage>. <publisher-loc>Antigua, Guatemala</publisher-loc>: <publisher-name>Centro de Investigaciones Regionales de Mesoam&#233;rica</publisher-name>.</mixed-citation></ref>
<ref id="B34"><label>34</label><mixed-citation publication-type="book"><string-name><surname>Chang</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Plauch&#233;</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Ohala</surname>, <given-names>J.</given-names></string-name> <year>2001</year>. <chapter-title>Markedness and consonant confusion asymme-tries</chapter-title>. In: <string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Hume</surname>, <given-names>E.</given-names></string-name> (eds.), <source>The role of speech perception in phonology</source>, <fpage>79</fpage>&#8211;<lpage>101</lpage>. <publisher-loc>New York, USA</publisher-loc>: <publisher-name>Academic Press</publisher-name>.</mixed-citation></ref>
<ref id="B35"><label>35</label><mixed-citation publication-type="book"><string-name><surname>Cojt&#237; Macario</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Lopez</surname>, <given-names>M.</given-names></string-name> <year>1990</year>. <chapter-title>Variaci&#243;n dialectal del idioma kaqchikel</chapter-title>. In: <string-name><surname>England</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Elliott</surname>, <given-names>S.</given-names></string-name> (eds.), <source>Lecturas sobre la ling&#252;&#237;stica maya</source>, <fpage>193</fpage>&#8211;<lpage>220</lpage>. <publisher-loc>Antigua, Guatemala</publisher-loc>: <publisher-name>Centro de Investigaciones Regionales de Mesoam&#233;rica</publisher-name>.</mixed-citation></ref>
<ref id="B36"><label>36</label><mixed-citation publication-type="journal"><string-name><surname>Coon</surname>, <given-names>J.</given-names></string-name> <year>2016</year>. <article-title>Mayan morphosyntax</article-title>. <source>Language and Linguistics Compass</source>, <volume>10</volume>(<issue>10</issue>), <fpage>515</fpage>&#8211;<lpage>550</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/lnc3.12149</pub-id></mixed-citation></ref>
<ref id="B37"><label>37</label><mixed-citation publication-type="journal"><string-name><surname>Cowan</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Morse</surname>, <given-names>P. A.</given-names></string-name> <year>1986</year>. <article-title>The use of auditory and phonetic memory in vowel discrimination</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>79</volume>(<issue>2</issue>), <fpage>500</fpage>&#8211;<lpage>507</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.393537</pub-id></mixed-citation></ref>
<ref id="B38"><label>38</label><mixed-citation publication-type="book"><string-name><surname>Cutler</surname>, <given-names>A.</given-names></string-name> <year>2012</year>. <source>Native listening: Language experience and the recognition of spoken words</source>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</mixed-citation></ref>
<ref id="B39"><label>39</label><mixed-citation publication-type="journal"><string-name><surname>Cutler</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Weber</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Smits</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Cooper</surname>, <given-names>N.</given-names></string-name> <year>2004</year>. <article-title>Patterns of English phoneme confusions by native and non-native listeners</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>116</volume>(<issue>6</issue>), <fpage>3668</fpage>&#8211;<lpage>3678</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1810292</pub-id></mixed-citation></ref>
<ref id="B40"><label>40</label><mixed-citation publication-type="journal"><string-name><surname>Daelemans</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Zavrel</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>van der Sloot</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>van den Bosch</surname>, <given-names>A.</given-names></string-name> <year>2009</year>. <source>TiMBL: Tilburg Memory-Based Learner</source> (Reference Guide No. Version 6.2). ILK Technical Report &#8211; ILK 09-01.</mixed-citation></ref>
<ref id="B41"><label>41</label><mixed-citation publication-type="journal"><string-name><surname>Dar</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Keren-Portnoy</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Vihman</surname>, <given-names>M.</given-names></string-name> <year>2018</year>. <article-title>An order effect in English infants&#8217; discrimination of an Urdu affricate contrast</article-title>. <source>Journal of Phonetics</source>, <volume>67</volume>, <fpage>49</fpage>&#8211;<lpage>64</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2017.12.002</pub-id></mixed-citation></ref>
<ref id="B42"><label>42</label><mixed-citation publication-type="journal"><string-name><surname>Davidson</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Martin</surname>, <given-names>A. E.</given-names></string-name> <year>2013</year>. <article-title>Modeling accuracy as a function of response time with the generalized linear mixed effects model</article-title>. <source>Acta psychologica</source>, <volume>144</volume>(<issue>1</issue>), <fpage>83</fpage>&#8211;<lpage>96</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.actpsy.2013.04.016</pub-id></mixed-citation></ref>
<ref id="B43"><label>43</label><mixed-citation publication-type="journal"><string-name><surname>Davidson</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Shaw</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Adams</surname>, <given-names>T.</given-names></string-name> <year>2007</year>. <article-title>The effect of word learning on the perception of non-native consonant sequences</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>122</volume>(<issue>6</issue>), <fpage>3697</fpage>&#8211;<lpage>3709</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.2801548</pub-id></mixed-citation></ref>
<ref id="B44"><label>44</label><mixed-citation publication-type="confproc"><string-name><surname>de Marneffe</surname>, <given-names>M.-C.</given-names></string-name>, <string-name><surname>Tomlinson</surname>, <given-names>J.</given-names>, <suffix>Jr.</suffix></string-name>, <string-name><surname>Tice</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Sumner</surname>, <given-names>M.</given-names></string-name> <year>2011</year>. <article-title>The interaction of lexical frequency and phonetic variation in the perception of accented speech</article-title>. In: <string-name><surname>Carlson</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Shipley</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Hoelscher</surname>, <given-names>C.</given-names></string-name> (eds.), <conf-name>The 33rd annual meeting of the Cognitive Science Society (CogSci 2011)</conf-name>, <fpage>3575</fpage>&#8211;<lpage>3580</lpage>. <conf-loc>New York, USA</conf-loc>: <conf-sponsor>Curran Associates, Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B45"><label>45</label><mixed-citation publication-type="journal"><string-name><surname>DiCanio</surname>, <given-names>C.</given-names></string-name> <year>2014</year>. <article-title>Cue weight in the perception of Trique glottal consonants</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>135</volume>(<issue>2</issue>), <fpage>884</fpage>&#8211;<lpage>895</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4861921</pub-id></mixed-citation></ref>
<ref id="B46"><label>46</label><mixed-citation publication-type="journal"><string-name><surname>DiCanio</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Nam</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Whalen</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Bunnell</surname>, <given-names>H. T.</given-names></string-name>, <string-name><surname>Amith</surname>, <given-names>J. D.</given-names></string-name>, &amp; <string-name><surname>Castillo Garc&#237;a</surname>, <given-names>R.</given-names></string-name> <year>2013</year>. <article-title>Using automatic alignment to analyze endangered language data: Testing the viability of untrained alignment</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>134</volume>(<issue>3</issue>), <fpage>2235</fpage>&#8211;<lpage>2246</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4816491</pub-id></mixed-citation></ref>
<ref id="B47"><label>47</label><mixed-citation publication-type="book"><string-name><surname>Dockum</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Campbell-Taylor</surname>, <given-names>E.</given-names></string-name> <year>2017</year>. <source>Minimum sufficient wordlist size for phonological typology</source>. Ms., <publisher-name>Yale University</publisher-name>.</mixed-citation></ref>
<ref id="B48"><label>48</label><mixed-citation publication-type="journal"><string-name><surname>Dubno</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Levitt</surname>, <given-names>H.</given-names></string-name> <year>1981</year>. <article-title>Predicting consonant confusions from acoustic analysis</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>69</volume>(<issue>1</issue>), <fpage>249</fpage>&#8211;<lpage>261</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.385345</pub-id></mixed-citation></ref>
<ref id="B49"><label>49</label><mixed-citation publication-type="thesis"><string-name><surname>DuBois</surname>, <given-names>J. W.</given-names></string-name> <year>1981</year>. <source>The Sacapultec language</source> (Unpublished doctoral dissertation). <publisher-name>University of California</publisher-name>, <publisher-loc>Berkeley</publisher-loc>.</mixed-citation></ref>
<ref id="B50"><label>50</label><mixed-citation publication-type="journal"><string-name><surname>Dunbar</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Idsardi</surname>, <given-names>W. J.</given-names></string-name> <year>2010</year>. <article-title>Review of Daniel Silverman (2006). A critical introduction to phonology: Of sound, mind, and body. London &amp; New York: Continuum. Pp. xii + 260</article-title>. <source>Phonology</source>, <volume>27</volume>(<issue>2</issue>), <fpage>325</fpage>&#8211;<lpage>331</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S095267571000014X</pub-id></mixed-citation></ref>
<ref id="B51"><label>51</label><mixed-citation publication-type="webpage"><string-name><surname>El Hattab</surname>, <given-names>H.</given-names></string-name> <year>2016</year>. <source>reveal.js</source>. <uri>https://github.com/hakimel/reveal.js/</uri>. <publisher-name>GitHub</publisher-name>.</mixed-citation></ref>
<ref id="B52"><label>52</label><mixed-citation publication-type="book"><string-name><surname>England</surname>, <given-names>N.</given-names></string-name> <year>1983</year>. <source>A grammar of Mam, a Mayan language</source>. <publisher-loc>Austin, Texas</publisher-loc>: <publisher-name>University of Texas Press</publisher-name>.</mixed-citation></ref>
<ref id="B53"><label>53</label><mixed-citation publication-type="book"><string-name><surname>England</surname>, <given-names>N.</given-names></string-name> <year>1996</year>. <chapter-title>The role of language standardization in revitalization</chapter-title>. In: <string-name><surname>Fischer</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Brown</surname>, <given-names>R. M.</given-names></string-name> (eds.), <source>Maya cultural activism in Guatemala</source>, <fpage>178</fpage>&#8211;<lpage>194</lpage>. <publisher-loc>Austin, Texas</publisher-loc>: <publisher-name>University of Texas Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1525/aa.2003.105.4.733</pub-id></mixed-citation></ref>
<ref id="B54"><label>54</label><mixed-citation publication-type="book"><string-name><surname>England</surname>, <given-names>N.</given-names></string-name> <year>2001</year>. <source>Introducci&#243;n a la gram&#225;tica de los idiomas mayas</source>. <publisher-loc>Ciudad de Guatemala, Guatemala</publisher-loc>: <publisher-name>Cholsamaj</publisher-name>.</mixed-citation></ref>
<ref id="B55"><label>55</label><mixed-citation publication-type="journal"><string-name><surname>England</surname>, <given-names>N.</given-names></string-name> <year>2003</year>. <article-title>Mayan language revival and revitalization politics: Linguists and linguistic ideologies</article-title>. <source>American Anthropologist</source>, <volume>105</volume>(<issue>4</issue>), <fpage>733</fpage>&#8211;<lpage>743</lpage>. DOI: <pub-id pub-id-type="doi">10.1525/aa.2003.105.4.733</pub-id></mixed-citation></ref>
<ref id="B56"><label>56</label><mixed-citation publication-type="journal"><string-name><surname>Ernestus</surname>, <given-names>M.</given-names></string-name> <year>2014</year>. <article-title>Acoustic reduction and the roles of abstractions and exemplars in speech processing</article-title>. <source>Lingua</source>, <volume>142</volume>, <fpage>27</fpage>&#8211;<lpage>41</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.lingua.2012.12.006</pub-id></mixed-citation></ref>
<ref id="B57"><label>57</label><mixed-citation publication-type="journal"><string-name><surname>Felty</surname>, <given-names>R. A.</given-names></string-name>, <string-name><surname>Buchwald</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Gruenenfelder</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Pisoni</surname>, <given-names>D.</given-names></string-name> <year>2013</year>. <article-title>Misperceptions of spoken words: Data from a random sample of American English words</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>134</volume>(<issue>1</issue>), <fpage>572</fpage>&#8211;<lpage>585</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4809540</pub-id></mixed-citation></ref>
<ref id="B58"><label>58</label><mixed-citation publication-type="journal"><string-name><surname>Ferrand</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>New</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Brysbaert</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Keuleers</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Bonin</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>M&#233;ot</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Pallier</surname>, <given-names>C.</given-names></string-name>, et al. <year>2010</year>. <article-title>The French lexicon project: Lexical decision data for 38,840 french words and 38,840 pseudowords</article-title>. <source>Behavior Research Methods</source>, <volume>42</volume>(<issue>2</issue>), <fpage>488</fpage>&#8211;<lpage>496</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BRM.42.2.488</pub-id></mixed-citation></ref>
<ref id="B59"><label>59</label><mixed-citation publication-type="book"><string-name><surname>Fischer</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Brown</surname>, <given-names>R. M.</given-names></string-name> (eds.) <year>1996</year>. <source>Maya cultural activism in Guatemala</source>. <publisher-loc>Austin, TX</publisher-loc>: <publisher-name>University of Texas Press</publisher-name>.</mixed-citation></ref>
<ref id="B60"><label>60</label><mixed-citation publication-type="journal"><string-name><surname>Fox</surname>, <given-names>R.</given-names></string-name> <year>1984</year>. <article-title>Effect of lexical status on phonetic categorization</article-title>. <source>Journal of Experimental Psychology: Human perception and performance</source>, <volume>10</volume>(<issue>4</issue>), <fpage>526</fpage>&#8211;<lpage>540</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0096-1523.10.4.526</pub-id></mixed-citation></ref>
<ref id="B61"><label>61</label><mixed-citation publication-type="book"><string-name><surname>Fre Woldu</surname>, <given-names>K.</given-names></string-name> <year>1985</year>. <source>The perception and production of Tigrinya stops</source>, <fpage>13</fpage>. <publisher-loc>Uppsala, Sweden</publisher-loc>: <publisher-name>Department of Linguistics, Uppsala University</publisher-name>.</mixed-citation></ref>
<ref id="B62"><label>62</label><mixed-citation publication-type="journal"><string-name><surname>Fujimura</surname>, <given-names>O.</given-names></string-name>, <string-name><surname>Macchi</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Streeter</surname>, <given-names>L.</given-names></string-name> <year>1978</year>. <article-title>Perception of stop consonants with conflicting transitional cues</article-title>. <source>Language and Speech</source>, <volume>21</volume>, <fpage>337</fpage>&#8211;<lpage>343</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/002383097802100408</pub-id></mixed-citation></ref>
<ref id="B63"><label>63</label><mixed-citation publication-type="journal"><string-name><surname>Gahl</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Yu</surname>, <given-names>A. C. L.</given-names></string-name> <year>2006</year>. <article-title>Introduction to the special issue on exemplar-based models in linguistics</article-title>. <source>The Linguistic Review</source>, <volume>23</volume>(<issue>3</issue>), <fpage>213</fpage>&#8211;<lpage>216</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/TLR.2006.007</pub-id></mixed-citation></ref>
<ref id="B64"><label>64</label><mixed-citation publication-type="thesis"><string-name><surname>Gallagher</surname>, <given-names>G.</given-names></string-name> <year>2010a</year>. <source>The perceptual basis of long-distance laryngeal restrictions</source> (Unpublished doctoral dissertation). <publisher-name>Massachusetts Institute of Technology</publisher-name>.</mixed-citation></ref>
<ref id="B65"><label>65</label><mixed-citation publication-type="journal"><string-name><surname>Gallagher</surname>, <given-names>G.</given-names></string-name> <year>2010b</year>. <article-title>Perceptual distinctness and long-distance laryngeal restrictions</article-title>. <source>Phonology</source>, <volume>27</volume>(<issue>3</issue>), <fpage>435</fpage>&#8211;<lpage>480</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0952675710000217</pub-id></mixed-citation></ref>
<ref id="B66"><label>66</label><mixed-citation publication-type="journal"><string-name><surname>Gallagher</surname>, <given-names>G.</given-names></string-name> <year>2011</year>. <article-title>Acoustic and articulatory features in phonology&#8211;the case for [long VOT]</article-title>. <source>The Linguistic Review</source>, <volume>28</volume>(<issue>3</issue>), <fpage>281</fpage>&#8211;<lpage>313</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/tlir.2011.008</pub-id></mixed-citation></ref>
<ref id="B67"><label>67</label><mixed-citation publication-type="journal"><string-name><surname>Gallagher</surname>, <given-names>G.</given-names></string-name> <year>2012</year>. <article-title>Perceptual similarity in non-local laryngeal restrictions</article-title>. <source>Lingua</source>, <volume>122</volume>(<issue>2</issue>), <fpage>112</fpage>&#8211;<lpage>124</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.lingua.2011.11.012</pub-id></mixed-citation></ref>
<ref id="B68"><label>68</label><mixed-citation publication-type="journal"><string-name><surname>Gallagher</surname>, <given-names>G.</given-names></string-name> <year>2014</year>. <article-title>An identity bias in phonotactics: Evidence from Cochabamba Quechua</article-title>. <source>Laboratory Phonology</source>, <volume>5</volume>(<issue>3</issue>), <fpage>337</fpage>&#8211;<lpage>378</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/lp-2014-0012</pub-id></mixed-citation></ref>
<ref id="B69"><label>69</label><mixed-citation publication-type="journal"><string-name><surname>Ganong</surname>, <given-names>W. F.</given-names></string-name> <year>1980</year>. <article-title>Phonetic categorization in auditory word perception</article-title>. <source>Journal of Experimental Psychology: Human Perception and Performance</source>, <volume>6</volume>(<issue>1</issue>), <fpage>110</fpage>&#8211;<lpage>125</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0096-1523.6.1.110</pub-id></mixed-citation></ref>
<ref id="B70"><label>70</label><mixed-citation publication-type="book"><string-name><surname>Garc&#237;a Matzar</surname>, <given-names>P. O.</given-names></string-name>, <string-name><surname>Toj Cotzajay</surname>, <given-names>V.</given-names></string-name>, &amp; <string-name><surname>Coc Tuiz</surname>, <given-names>D.</given-names></string-name> <year>1999</year>. <source>Gram&#225;tica del idioma Kaqchikel</source>. <publisher-loc>Antigua, Guatemala</publisher-loc>: <publisher-name>Proyecto Ling&#252;&#237;stico Francisco Marroqu&#237;n</publisher-name>.</mixed-citation></ref>
<ref id="B71"><label>71</label><mixed-citation publication-type="confproc"><string-name><surname>Gasser</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Bowern</surname>, <given-names>C.</given-names></string-name> <year>2014</year>. <article-title>Revisiting phonotactic generalizations in Australian languages</article-title>. In: <string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Moore-Cantwell</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Pater</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Staubs</surname>, <given-names>R.</given-names></string-name> (eds.), <conf-name>Proceedings of the 2013 annual meetings on phonology</conf-name>. DOI: <pub-id pub-id-type="doi">10.3765/amp.v1i1.17</pub-id></mixed-citation></ref>
<ref id="B72"><label>72</label><mixed-citation publication-type="journal"><string-name><surname>Goldinger</surname>, <given-names>S. D.</given-names></string-name> <year>1996</year>. <article-title>Words and voices: Episodic traces in spoken word identification and recognition memory</article-title>. <source>Journal of Experimental Psychology: Learning, Memory, and Cognition</source>, <volume>22</volume>(<issue>5</issue>), <fpage>1166</fpage>&#8211;<lpage>1183</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0278-7393.22.5.1166</pub-id></mixed-citation></ref>
<ref id="B73"><label>73</label><mixed-citation publication-type="journal"><string-name><surname>Goldinger</surname>, <given-names>S. D.</given-names></string-name> <year>1998</year>. <article-title>Echoes of echoes? An episodic theory of lexical access</article-title>. <source>Psychological review</source>, <volume>105</volume>(<issue>2</issue>), <fpage>251</fpage>&#8211;<lpage>279</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0033-295X.105.2.251</pub-id></mixed-citation></ref>
<ref id="B74"><label>74</label><mixed-citation publication-type="journal"><string-name><surname>Goldrick</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Vaughn</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Murphy</surname>, <given-names>A.</given-names></string-name> <year>2013</year>. <article-title>The effects of lexical neighbors on stop consonant articulation</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>134</volume>(<issue>2</issue>), <fpage>EL172</fpage>&#8211;<lpage>EL177</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4812821</pub-id></mixed-citation></ref>
<ref id="B75"><label>75</label><mixed-citation publication-type="webpage"><string-name><surname>Gorman</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Howell</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Wagner</surname>, <given-names>M.</given-names></string-name> <year>2011</year>. <article-title>Prosodylab-aligner: A tool for forced alignment of laboratory speech</article-title>. <source>Canadian Acoustics</source>, <volume>39</volume>(<issue>3</issue>), <fpage>192</fpage>&#8211;<lpage>193</lpage>. Retrieved from: <uri>https://jcaa.caa-aca.ca/index.php/jcaa/article/view/2476</uri>.</mixed-citation></ref>
<ref id="B76"><label>76</label><mixed-citation publication-type="thesis"><string-name><surname>Graff</surname>, <given-names>P.</given-names></string-name> <year>2012</year>. <source>Communicative efficiency in the lexicon</source> (Unpublished doctoral dissertation). <publisher-name>Massachusetts Institute of Technology</publisher-name>.</mixed-citation></ref>
<ref id="B77"><label>77</label><mixed-citation publication-type="thesis"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name> <year>2009</year>. <source>A probabilistic model of phonological relationships from contrast to allophony</source> (Unpublished doctoral dissertation). <publisher-name>The Ohio State University</publisher-name>.</mixed-citation></ref>
<ref id="B78"><label>78</label><mixed-citation publication-type="journal"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name> <year>2012</year>. <article-title>Phonological relationships: A probabilistic model</article-title>. <source>McGill Working Papers in Linguistics</source>, <volume>22</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>14</lpage>.</mixed-citation></ref>
<ref id="B79"><label>79</label><mixed-citation publication-type="journal"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name> <year>2013</year>. <article-title>A typology of intermediate phonological relationships</article-title>. <source>The Linguistic Review</source>, <volume>30</volume>(<issue>2</issue>), <fpage>215</fpage>&#8211;<lpage>275</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/tlr-2013-0008</pub-id></mixed-citation></ref>
<ref id="B80"><label>80</label><mixed-citation publication-type="webpage"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name>, <string-name><surname>Allen</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Fry</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Mackie</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>McAuliffe</surname>, <given-names>M.</given-names></string-name> <year>2015</year>. <source>Phonological CorpusTools, Version 1.1. [Computer program]</source>. <uri>http://phonologicalcorpustools.github.io/CorpusTools/</uri>. <publisher-name>Github</publisher-name>.</mixed-citation></ref>
<ref id="B81"><label>81</label><mixed-citation publication-type="journal"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name>, &amp; <string-name><surname>Hume</surname>, <given-names>E.</given-names></string-name> (in preparation). <article-title>Modeling perceived similarity: The influence of phonetics, phonology and frequency on the perception of French vowels</article-title>. <source>Laboratory Phonology</source>.</mixed-citation></ref>
<ref id="B82"><label>82</label><mixed-citation publication-type="journal"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name>, <string-name><surname>Hume</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Jaeger</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Wedel</surname>, <given-names>A.</given-names></string-name> (in preparation). <article-title>The message shapes phonology</article-title>.</mixed-citation></ref>
<ref id="B83"><label>83</label><mixed-citation publication-type="confproc"><string-name><surname>Hall</surname>, <given-names>K. C.</given-names></string-name>, <string-name><surname>Letawsky</surname>, <given-names>V.</given-names></string-name>, <string-name><surname>Turner</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Allen</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>McMullin</surname>, <given-names>K.</given-names></string-name> <year>2014</year>. <article-title>Effects of predictability of distribution on within-language perception</article-title>. In: <string-name><surname>V&#299;nerte</surname>, <given-names>S.</given-names></string-name> (ed.), <conf-name>Proceedings of the 2015 annual conference of the Canadian Linguistics Association</conf-name>, <fpage>1</fpage>&#8211;<lpage>15</lpage>. <conf-loc>Ottawa, Canada</conf-loc>: <conf-sponsor>Canadian Linguistics Association</conf-sponsor>. (Available online at: <uri>http://cla-acl.ca/wp-content/uploads/Hall_Letawsky_Turner_Allen_McMullin-2015.pdf</uri>).</mixed-citation></ref>
<ref id="B84"><label>84</label><mixed-citation publication-type="journal"><string-name><surname>Harnsberger</surname>, <given-names>J.</given-names></string-name> <year>2000</year>. <article-title>A cross-language study of the identification of non-native nasal consonants varying in place of articulation</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>108</volume>(<issue>2</issue>), <fpage>764</fpage>&#8211;<lpage>783</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.429610</pub-id></mixed-citation></ref>
<ref id="B85"><label>85</label><mixed-citation publication-type="journal"><string-name><surname>Harnsberger</surname>, <given-names>J.</given-names></string-name> <year>2001a</year>. <article-title>On the relationship between identification and discrimination of non-native nasal consonants</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>110</volume>(<issue>1</issue>), <fpage>489</fpage>&#8211;<lpage>503</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1371758</pub-id></mixed-citation></ref>
<ref id="B86"><label>86</label><mixed-citation publication-type="journal"><string-name><surname>Harnsberger</surname>, <given-names>J.</given-names></string-name> <year>2001b</year>. <article-title>The perception of Malayalam nasal consonants by Marathi, Punjabi, Tamil, Oriya, Bengali, and American English listeners: A multidimensional scaling analysis</article-title>. <source>Journal of Phonetics</source>, <volume>29</volume>(<issue>3</issue>), <fpage>303</fpage>&#8211;<lpage>327</lpage>. DOI: <pub-id pub-id-type="doi">10.1006/jpho.2001.0140</pub-id></mixed-citation></ref>
<ref id="B87"><label>87</label><mixed-citation publication-type="book"><string-name><surname>Harris</surname>, <given-names>J.</given-names></string-name> <year>1969</year>. <source>Spanish phonology</source>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</mixed-citation></ref>
<ref id="B88"><label>88</label><mixed-citation publication-type="journal"><string-name><surname>Hayes</surname>, <given-names>B.</given-names></string-name>, &amp; <string-name><surname>Wilson</surname>, <given-names>C.</given-names></string-name> <year>2008</year>. <article-title>A maximum entropy model of phonotactics and phonotactic learning</article-title>. <source>Linguistic Inquiry</source>, <volume>39</volume>(<issue>3</issue>), <fpage>379</fpage>&#8211;<lpage>440</lpage>. DOI: <pub-id pub-id-type="doi">10.1162/ling.2008.39.3.379</pub-id></mixed-citation></ref>
<ref id="B89"><label>89</label><mixed-citation publication-type="journal"><string-name><surname>Heitz</surname>, <given-names>R. P.</given-names></string-name> <year>2014</year>. <article-title>The speed-accuracy tradeoff: History, physiology, methodology, and behavior</article-title>. <source>Frontiers in neuroscience</source>, <volume>8</volume>, <fpage>1</fpage>&#8211;<lpage>9</lpage>. (Article 150). DOI: <pub-id pub-id-type="doi">10.3389/fnins.2014.00150</pub-id></mixed-citation></ref>
<ref id="B90"><label>90</label><mixed-citation publication-type="journal"><string-name><surname>Henrich</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Heine</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Norenzayan</surname>, <given-names>A.</given-names></string-name> <year>2010</year>. <article-title>The weirdest people in the world?</article-title> <source>Behavioral and brain sciences</source>, <volume>33</volume>(<issue>2&#8211;3</issue>), <fpage>61</fpage>&#8211;<lpage>83</lpage>. DOI: <pub-id pub-id-type="doi">10.2139/ssrn.1601785</pub-id></mixed-citation></ref>
<ref id="B91"><label>91</label><mixed-citation publication-type="journal"><string-name><surname>Hockett</surname>, <given-names>C.</given-names></string-name> <year>1967</year>. <article-title>The quantification of functional load</article-title>. <source>Word</source>, <volume>23</volume>(<issue>1&#8211;3</issue>), <fpage>300</fpage>&#8211;<lpage>320</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/00437956.1967.11435484</pub-id></mixed-citation></ref>
<ref id="B92"><label>92</label><mixed-citation publication-type="journal"><string-name><surname>Holt</surname>, <given-names>L.</given-names></string-name>, &amp; <string-name><surname>Lotto</surname>, <given-names>A.</given-names></string-name> <year>2006</year>. <article-title>Cue weighting in auditory categorization: Implications for first and second language acquisition</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>119</volume>(<issue>5</issue>), <fpage>3059</fpage>&#8211;<lpage>3071</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.2188377</pub-id></mixed-citation></ref>
<ref id="B93"><label>93</label><mixed-citation publication-type="journal"><string-name><surname>Howes</surname>, <given-names>D.</given-names></string-name> <year>1957</year>. <article-title>On the relation between the intelligibility and frequency of occurrence of English words</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>29</volume>(<issue>2</issue>), <fpage>296</fpage>&#8211;<lpage>305</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1908862</pub-id></mixed-citation></ref>
<ref id="B94"><label>94</label><mixed-citation publication-type="webpage"><string-name><surname>Hyman</surname>, <given-names>L.</given-names></string-name> <year>2015</year>. <chapter-title>Why underlying representations?</chapter-title> In: <source>UC Berkeley Phonology Lab annual report</source>, <fpage>210</fpage>&#8211;<lpage>226</lpage>. <publisher-name>Department of Linguistics, UC Berkeley</publisher-name>. (Available online at: <uri>http://escholarship.org/uc/item/7hn3623c</uri>).</mixed-citation></ref>
<ref id="B95"><label>95</label><mixed-citation publication-type="book"><string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> <year>2005</year>. <chapter-title>Speaker normalization in speech perception</chapter-title>. In: <string-name><surname>Pisoni</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Remez</surname>, <given-names>R.</given-names></string-name> (eds.), <source>The handbook of speech perception</source>, <fpage>363</fpage>&#8211;<lpage>389</lpage>. <publisher-loc>Malden, MA</publisher-loc>: <publisher-name>Blackwell</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1002/9780470757024.ch15</pub-id></mixed-citation></ref>
<ref id="B96"><label>96</label><mixed-citation publication-type="journal"><string-name><surname>Johnson</surname>, <given-names>P. C.</given-names></string-name> <year>2014</year>. <article-title>Extension of Nakagawa &amp; Schielzeth&#8217;s R2GLMM to random slopes models</article-title>. <source>Methods in Ecology and Evolution</source>, n/a&#8211;n/a. DOI: <pub-id pub-id-type="doi">10.1111/2041-210X.12225</pub-id></mixed-citation></ref>
<ref id="B97"><label>97</label><mixed-citation publication-type="book"><string-name><surname>Jun</surname>, <given-names>J.</given-names></string-name> <year>2004</year>. <chapter-title>Place assimilation</chapter-title>. In: <string-name><surname>Hayes</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Kirchner</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Steriade</surname>, <given-names>D.</given-names></string-name> (eds.), <source>Phonetically based phonology</source>, <fpage>58</fpage>&#8211;<lpage>86</lpage>. <publisher-loc>Cambridge, UK</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1017/CBO9780511486401.003</pub-id></mixed-citation></ref>
<ref id="B98"><label>98</label><mixed-citation publication-type="journal"><string-name><surname>Jurafsky</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Bell</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Gregory</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Raymond</surname>, <given-names>W. D.</given-names></string-name> <year>2001</year>. <article-title>Probabilistic relations between words: Evidence from reduction in lexical production</article-title>. <source>Typological studies in language</source>, <volume>45</volume>, <fpage>229</fpage>&#8211;<lpage>254</lpage>. DOI: <pub-id pub-id-type="doi">10.1075/tsl.45.13jur</pub-id></mixed-citation></ref>
<ref id="B99"><label>99</label><mixed-citation publication-type="webpage"><string-name><surname>Kataoka</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> <year>2007</year>. <chapter-title>Frequency effects in cross-linguistic stop place perception: A case of /t/ &#8211; /k/ in Japanese and English</chapter-title>. In: <source>UC Berkeley Phonology Lab annual report</source>, <fpage>273</fpage>&#8211;<lpage>301</lpage>. <publisher-name>Department of Linguistics, UC Berkeley</publisher-name>. (Available online at: <uri>http://linguistics.berkeley.edu/phonlab/documents/2007/Kataoka_Johnson.pdf</uri>).</mixed-citation></ref>
<ref id="B100"><label>100</label><mixed-citation publication-type="book"><string-name><surname>Kaufman</surname>, <given-names>T.</given-names></string-name> <year>1990</year>. <chapter-title>Algunos rasgos estructurales de los idiomas mayances con referencia especial al K&#8217;iche&#8217;</chapter-title>. In: <string-name><surname>England</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Elliott</surname>, <given-names>S.</given-names></string-name> (eds.), <source>Lecturas sobre la ling&#252;&#237;stica maya</source>, <fpage>59</fpage>&#8211;<lpage>114</lpage>. <publisher-loc>Antigua, Guatemala</publisher-loc>: <publisher-name>Centro de Investigaciones Regionales de Mesoam&#233;rica</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1086/466206</pub-id></mixed-citation></ref>
<ref id="B101"><label>101</label><mixed-citation publication-type="webpage"><string-name><surname>Kaufman</surname>, <given-names>T.</given-names></string-name> <year>2003</year>. <source>A preliminary Mayan etymological dictionary</source>. Ms., <publisher-name>Foundation for the Advancement of Mesoamerican Studies</publisher-name>. Available online at: <uri>http://www.famsi.org/reports/01051/</uri>.</mixed-citation></ref>
<ref id="B102"><label>102</label><mixed-citation publication-type="webpage"><string-name><surname>Keuleers</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Diependaele</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Brysbaert</surname>, <given-names>M.</given-names></string-name> <year>2010</year>. <article-title>Practice effects in large-scale visual word recognition studies: A lexical decision study on 14,000 dutch mono- and disyllabic words and nonwords</article-title>. <source>Frontiers in Psychology</source>, <volume>1</volume>, <fpage>174</fpage>. Retrieved from: <uri>http://journal.frontiersin.org/article/10.3389/fpsyg.2010.00174</uri>. DOI: <pub-id pub-id-type="doi">10.3389/fpsyg.2010.00174</pub-id></mixed-citation></ref>
<ref id="B103"><label>103</label><mixed-citation publication-type="journal"><string-name><surname>Keuleers</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Lacey</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Rastle</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Brysbaert</surname>, <given-names>M.</given-names></string-name> <year>2012</year>. <article-title>The British Lexicon Project: Lexical decision data for 28,730 monosyllabic and disyllabic English words</article-title>. <source>Behavior Research Methods</source>, <volume>44</volume>(<issue>1</issue>), <fpage>287</fpage>&#8211;<lpage>304</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/s13428-011-0118-4</pub-id></mixed-citation></ref>
<ref id="B104"><label>104</label><mixed-citation publication-type="journal"><string-name><surname>King</surname>, <given-names>R. D.</given-names></string-name> <year>1967</year>. <article-title>Functional load and sound change</article-title>. <source>Language</source>, <fpage>831</fpage>&#8211;<lpage>852</lpage>. DOI: <pub-id pub-id-type="doi">10.2307/411969</pub-id></mixed-citation></ref>
<ref id="B105"><label>105</label><mixed-citation publication-type="thesis"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name> <year>1984</year>. <source>The phonetics and phonology of the timing of oral and glottal events</source> (Unpublished doctoral dissertation). <publisher-name>University of California</publisher-name>, <publisher-loc>Berkeley</publisher-loc>.</mixed-citation></ref>
<ref id="B106"><label>106</label><mixed-citation publication-type="book"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name> <year>2005a</year>. <chapter-title>Ears to categories: New arguments for autonomy</chapter-title>. In: <string-name><surname>Frota</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Vig&#225;rio</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Freitas</surname>, <given-names>M.</given-names></string-name> (eds.), <source>Prosodies: with special reference to Iberian languages</source>, <fpage>177</fpage>&#8211;<lpage>222</lpage>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1515/9783110197587.2.177</pub-id></mixed-citation></ref>
<ref id="B107"><label>107</label><mixed-citation publication-type="book"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name> <year>2005b</year>. <chapter-title>The phonetics of Athabaskan tonogenesis</chapter-title>. In: <string-name><surname>Hargus</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Rice</surname>, <given-names>K.</given-names></string-name> (eds.), <source>Athabaskan prosody</source>, <fpage>137</fpage>&#8211;<lpage>184</lpage>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1075/cilt.269.09kin</pub-id></mixed-citation></ref>
<ref id="B108"><label>108</label><mixed-citation publication-type="journal"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Levy</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Rysling</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Staub</surname>, <given-names>A.</given-names></string-name> <year>2016</year>. <article-title>Eye movement evidence for an immediate Ganong effect</article-title>. <source>Journal of experimental psychology: Human perception and performance</source>, <volume>42</volume>(<issue>12</issue>), <fpage>1969</fpage>&#8211;<lpage>1988</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/xhp0000269</pub-id></mixed-citation></ref>
<ref id="B109"><label>109</label><mixed-citation publication-type="book"><string-name><surname>Kuhl</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Iverson</surname>, <given-names>P.</given-names></string-name> <year>1995</year>. <chapter-title>Linguistic experience and the &#8220;perceptual magnet effect&#8221;</chapter-title>. In: <string-name><surname>Strange</surname>, <given-names>W.</given-names></string-name> (ed.), <source>Speech perception and linguistic experience: issues in cross-language research</source>, <fpage>121</fpage>&#8211;<lpage>154</lpage>. <publisher-loc>Baltimore, MD</publisher-loc>: <publisher-name>York Press</publisher-name>.</mixed-citation></ref>
<ref id="B110"><label>110</label><mixed-citation publication-type="journal"><string-name><surname>Kullback</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Leibler</surname>, <given-names>R.</given-names></string-name> <year>1951</year>. <article-title>On information and sufficiency</article-title>. <source>The Annals of Mathematical Statistics</source>, <volume>22</volume>(<issue>1</issue>), <fpage>79</fpage>&#8211;<lpage>86</lpage>. DOI: <pub-id pub-id-type="doi">10.1214/aoms/1177729694</pub-id></mixed-citation></ref>
<ref id="B111"><label>111</label><mixed-citation publication-type="book"><string-name><surname>Ku&#269;era</surname>, <given-names>H.</given-names></string-name> (<year>1963</year>). <source>Entropy, redundancy and functional load in Russian and Czech</source>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton &amp; Company</publisher-name>.</mixed-citation></ref>
<ref id="B112"><label>112</label><mixed-citation publication-type="book"><string-name><surname>Ladefoged</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Disner</surname>, <given-names>S. F.</given-names></string-name> <year>2012</year>. <source>Vowels and consonants</source> (<edition>3rd ed.</edition>). <publisher-loc>Malden, MA</publisher-loc>: <publisher-name>Wiley-Blackwell</publisher-name>.</mixed-citation></ref>
<ref id="B113"><label>113</label><mixed-citation publication-type="thesis"><string-name><surname>Larsen</surname>, <given-names>T.</given-names></string-name> <year>1988</year>. <source>Manifestations of ergativity in Quich&#233; grammar</source> (Unpublished doctoral dissertation). <publisher-name>University of California</publisher-name>, <publisher-loc>Berkeley</publisher-loc>.</mixed-citation></ref>
<ref id="B114"><label>114</label><mixed-citation publication-type="journal"><string-name><surname>Lindau</surname>, <given-names>M.</given-names></string-name> <year>1984</year>. <article-title>Phonetic differences in glottalic consonants</article-title>. <source>Journal of Phonetics</source>, <volume>12</volume>, <fpage>147</fpage>&#8211;<lpage>155</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.2019283</pub-id></mixed-citation></ref>
<ref id="B115"><label>115</label><mixed-citation publication-type="book"><string-name><surname>Lodge</surname>, <given-names>K.</given-names></string-name> <year>2009</year>. <source>Fundamental concepts in phonology: Sameness and difference</source>. <publisher-loc>Edinburgh</publisher-loc>: <publisher-name>Edinburgh University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.3366/edinburgh/9780748625659.001.0001</pub-id></mixed-citation></ref>
<ref id="B116"><label>116</label><mixed-citation publication-type="book"><string-name><surname>Macario</surname>, <given-names>N. C.</given-names></string-name>, <string-name><surname>Cutzal</surname>, <given-names>M. C.</given-names></string-name>, &amp; <string-name><surname>Semey&#225;</surname>, <given-names>M. A. C.</given-names></string-name> <year>1998</year>. <source>Diccionario Kaqchikel</source>. <publisher-loc>Antigua, Guatemala</publisher-loc>: <publisher-name>Proyecto Ling&#252;&#237;stico Francisco Marroqu&#237;n</publisher-name>.</mixed-citation></ref>
<ref id="B117"><label>117</label><mixed-citation publication-type="confproc"><string-name><surname>Macklin-Cordes</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Round</surname>, <given-names>E.</given-names></string-name> <year>2015</year>, <conf-date>4&#8211;6 November 2015</conf-date>. <article-title>High-definition phonotactics reflect linguistic pasts</article-title>. In: <string-name><surname>Wahle</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Kollner</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Baayen</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Jager</surname>, <given-names>G.</given-names></string-name>, &amp; <string-name><surname>Baayen-Oudshoorn</surname>, <given-names>T.</given-names></string-name> (eds.), <conf-name>Proceedings of the 6th conference on quantitative investigations in theoretical linguistics</conf-name>. <conf-loc>Tubingen, Germany</conf-loc>. DOI: <pub-id pub-id-type="doi">10.15496/publikation-8609</pub-id></mixed-citation></ref>
<ref id="B118"><label>118</label><mixed-citation publication-type="book"><string-name><surname>Macmillan</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Creelman</surname>, <given-names>C. D.</given-names></string-name> <year>2005</year>. <source>Detection theory: A user&#8217;s guide</source> (<edition>2nd ed.</edition>). <publisher-loc>Mahwah, NJ</publisher-loc>: <publisher-name>Lawrence Erlbaum Associates</publisher-name>. DOI: <pub-id pub-id-type="doi">10.4324/9781410611147</pub-id></mixed-citation></ref>
<ref id="B119"><label>119</label><mixed-citation publication-type="webpage"><string-name><surname>Maddieson</surname>, <given-names>I.</given-names></string-name> <year>2009</year>. <chapter-title>Glottalized consonants</chapter-title>. <source>The World Atlas of Language Structures Online (WALS)</source>. <publisher-loc>Munich</publisher-loc>: <publisher-name>Max Planck Digital Library</publisher-name>. (Available online at: <uri>http://wals.info/feature/7</uri>).</mixed-citation></ref>
<ref id="B120"><label>120</label><mixed-citation publication-type="webpage"><string-name><surname>Maekawa</surname>, <given-names>K.</given-names></string-name> <year>2003</year>. <chapter-title>Corpus of Spontaneous Japanese: Its design and evaluation</chapter-title>. In: <source>Spontaneous speech processing and recognition (SSPR 2003)</source>. <publisher-name>ICSA Speech Archive</publisher-name>. (<uri>http://www.isca-speech.org/archive_open/sspr2003/sspr_mmo2.html</uri>).</mixed-citation></ref>
<ref id="B121"><label>121</label><mixed-citation publication-type="book"><string-name><surname>Majzul</surname>, <given-names>L. F. P.</given-names></string-name> <year>2007</year>. <source>Rusoltzij ri Kaqchikel: Diccionario est&#225;ndar biling&#252;e Kaqchikel-Espa&#241;ol</source>. <publisher-loc>Ciudad de Guatemala, Guatemala</publisher-loc>: <publisher-name>Cholsamaj</publisher-name>.</mixed-citation></ref>
<ref id="B122"><label>122</label><mixed-citation publication-type="book"><string-name><surname>Majzul</surname>, <given-names>L. F. P.</given-names></string-name>, <string-name><surname>Matzar</surname>, <given-names>P. O. G.</given-names></string-name>, &amp; <string-name><surname>Serech</surname>, <given-names>C. E.</given-names></string-name> <year>2000</year>. <source>Rujunamaxik ri Kaqchikel chi&#8217;: Variaci&#243;n dialectal en Kaqchikel</source>. <publisher-loc>Ciudad de Guatemala, Guatemala</publisher-loc>: <publisher-name>Cholsamaj</publisher-name>.</mixed-citation></ref>
<ref id="B123"><label>123</label><mixed-citation publication-type="journal"><string-name><surname>Martinet</surname>, <given-names>A.</given-names></string-name> <year>1952</year>. <article-title>Function, structure, and sound change</article-title>. <source>Word</source>, <volume>8</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>32</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/00437956.1952.11659416</pub-id></mixed-citation></ref>
<ref id="B124"><label>124</label><mixed-citation publication-type="book"><string-name><surname>Maxwell</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Hill</surname>, <given-names>R.</given-names></string-name> <year>2010</year>. <source>Kaqchikel chronicles: the definitive edition</source>. <publisher-loc>Austin, TX</publisher-loc>: <publisher-name>University of Texas Press</publisher-name>.</mixed-citation></ref>
<ref id="B125"><label>125</label><mixed-citation publication-type="journal"><string-name><surname>McClelland</surname>, <given-names>J. L.</given-names></string-name>, &amp; <string-name><surname>Elman</surname>, <given-names>J. L.</given-names></string-name> <year>1986</year>. <article-title>The TRACE model of speech perception</article-title>. <source>Cognitive Psychology</source>, <volume>18</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>86</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/0010-0285(86)90015-0</pub-id></mixed-citation></ref>
<ref id="B126"><label>126</label><mixed-citation publication-type="journal"><string-name><surname>McClelland</surname>, <given-names>J. L.</given-names></string-name>, <string-name><surname>Mirman</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Holt</surname>, <given-names>L.</given-names></string-name> <year>2006</year>. <article-title>Are there interactive processes in speech perception?</article-title> <source>Trends in cognitive sciences</source>, <volume>10</volume>(<issue>8</issue>), <fpage>363</fpage>&#8211;<lpage>369</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.tics.2006.06.007</pub-id></mixed-citation></ref>
<ref id="B127"><label>127</label><mixed-citation publication-type="book"><string-name><surname>McClelland</surname>, <given-names>J. L.</given-names></string-name>, <string-name><surname>Rumelhart</surname>, <given-names>D. E.</given-names></string-name>, &amp; <string-name><surname>Hinton</surname>, <given-names>G. E.</given-names></string-name> <year>1986</year>. <chapter-title>The appeal of parallel distributed processing</chapter-title>. In: <string-name><surname>Rumelhart</surname>, <given-names>D. E.</given-names></string-name>, <string-name><surname>McClelland</surname>, <given-names>J. L.</given-names></string-name>, &amp; <collab>The PDP Research Group</collab>. (eds.), <source>Parallel distributed processing: Explorations in the microstructure of cognition</source>, <volume>1</volume>, <fpage>3</fpage>&#8211;<lpage>44</lpage>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1016/B978-1-4832-1446-7.50010-8</pub-id></mixed-citation></ref>
<ref id="B128"><label>128</label><mixed-citation publication-type="webpage"><string-name><surname>McCloy</surname>, <given-names>D.</given-names></string-name> <year>2014</year>. <source>praat-semiauto</source>. <uri>https://github.com/drammock/praat-semiauto/</uri>. <publisher-name>GitHub</publisher-name>.</mixed-citation></ref>
<ref id="B129"><label>129</label><mixed-citation publication-type="thesis"><string-name><surname>McGuire</surname>, <given-names>G.</given-names></string-name> <year>2007</year>. <source>Phonetic category learning</source> (Unpublished doctoral dissertation). <publisher-name>The Ohio State University</publisher-name>.</mixed-citation></ref>
<ref id="B130"><label>130</label><mixed-citation publication-type="webpage"><string-name><surname>McGuire</surname>, <given-names>G.</given-names></string-name> <year>2010</year>. <source>A brief primer on experimental designs for speech perception research</source>. Ms. <publisher-name>University of California</publisher-name>, <publisher-loc>Santa Cruz</publisher-loc>. (Available online at: <uri>http://people.ucsc.edu/gmcguir1/experiment_designs.pdf</uri>).</mixed-citation></ref>
<ref id="B131"><label>131</label><mixed-citation publication-type="journal"><string-name><surname>Meyer</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Dentel</surname>, <given-names>L.</given-names></string-name>, &amp; <string-name><surname>Meunier</surname>, <given-names>F.</given-names></string-name> <year>2013</year>, 11. <article-title>Speech recognition in natural background noise</article-title>. <source>PLoS ONE</source>, <volume>8</volume>(<issue>11</issue>), <fpage>1</fpage>&#8211;<lpage>14</lpage>. DOI: <pub-id pub-id-type="doi">10.1371/journal.pone.0079279</pub-id></mixed-citation></ref>
<ref id="B132"><label>132</label><mixed-citation publication-type="journal"><string-name><surname>Mielke</surname>, <given-names>J.</given-names></string-name> <year>2012</year>. <article-title>A phonetically-based metric of sound similarity</article-title>. <source>Lingua</source>, <volume>122</volume>, <fpage>145</fpage>&#8211;<lpage>163</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.lingua.2011.04.006</pub-id></mixed-citation></ref>
<ref id="B133"><label>133</label><mixed-citation publication-type="journal"><string-name><surname>Nakagawa</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Schielzeth</surname>, <given-names>H.</given-names></string-name> <year>2013</year>. <article-title>A general and simple method for obtaining r2 from generalized linear mixed-effects models</article-title>. <source>Methods in Ecology and Evolution</source>, <volume>4</volume>(<issue>2</issue>), <fpage>133</fpage>&#8211;<lpage>142</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/j.2041-210x.2012.00261.x</pub-id></mixed-citation></ref>
<ref id="B134"><label>134</label><mixed-citation publication-type="journal"><string-name><surname>Nelson</surname>, <given-names>N. R.</given-names></string-name>, &amp; <string-name><surname>Wedel</surname>, <given-names>A.</given-names></string-name> <year>2017</year>. <article-title>The phonetic specificity of competition: Contrastive hyperarticulation of voice onset time in conversational English</article-title>. <source>Journal of Phonetics</source>, <volume>64</volume>, <fpage>51</fpage>&#8211;<lpage>70</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2017.01.008</pub-id></mixed-citation></ref>
<ref id="B135"><label>135</label><mixed-citation publication-type="journal"><string-name><surname>Norris</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>McQueen</surname>, <given-names>J. M.</given-names></string-name>, &amp; <string-name><surname>Cutler</surname>, <given-names>A.</given-names></string-name> <year>2000</year>. <article-title>Merging information in speech recognition: Feedback is never necessary</article-title>. <source>Behavioral and Brain Sciences</source>, <volume>23</volume>(<issue>3</issue>), <fpage>299</fpage>&#8211;<lpage>325</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0140525X00003241</pub-id></mixed-citation></ref>
<ref id="B136"><label>136</label><mixed-citation publication-type="journal"><string-name><surname>Oh</surname>, <given-names>Y. M.</given-names></string-name>, <string-name><surname>Coup&#233;</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Marsico</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Pellegrino</surname>, <given-names>F.</given-names></string-name> <year>2015</year>. <article-title>Bridging phonological system and lexicon: Insights from a corpus study of functional load</article-title>. <source>Journal of phonetics</source>, <volume>53</volume>, <fpage>153</fpage>&#8211;<lpage>176</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2015.08.003</pub-id></mixed-citation></ref>
<ref id="B137"><label>137</label><mixed-citation publication-type="journal"><string-name><surname>Oh</surname>, <given-names>Y. M.</given-names></string-name>, <string-name><surname>Pellegrino</surname>, <given-names>F.</given-names></string-name>, <string-name><surname>Coup&#233;</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Marsico</surname>, <given-names>E.</given-names></string-name> <year>2013</year>. <article-title>Cross-language comparison of functional load for vowels, consonants, and tones</article-title>. In: <source>Proceedings of interspeech</source>, <fpage>3032</fpage>&#8211;<lpage>3036</lpage>.</mixed-citation></ref>
<ref id="B138"><label>138</label><mixed-citation publication-type="book"><string-name><surname>Ohala</surname>, <given-names>J.</given-names></string-name> <year>1993</year>. <chapter-title>The phonetics of sound change</chapter-title>. In: <string-name><surname>Jones</surname>, <given-names>C.</given-names></string-name> (ed.), <source>Historical linguistics: Problems and perspectives</source>, <fpage>237</fpage>&#8211;<lpage>278</lpage>. <publisher-loc>London</publisher-loc>: <publisher-name>Longman</publisher-name>.</mixed-citation></ref>
<ref id="B139"><label>139</label><mixed-citation publication-type="journal"><string-name><surname>Peirce</surname>, <given-names>J. W.</given-names></string-name> <year>2007</year>. <article-title>Psychopy&#8211;psychophysics software in Python</article-title>. <source>Journal of Neuroscience Methods</source>, <volume>162</volume>(<issue>1&#8211;2</issue>), <fpage>8</fpage>&#8211;<lpage>13</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.jneumeth.2006.11.017</pub-id></mixed-citation></ref>
<ref id="B140"><label>140</label><mixed-citation publication-type="journal"><string-name><surname>Peperkamp</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Le Calvez</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Nadal</surname>, <given-names>J.-P.</given-names></string-name>, &amp; <string-name><surname>Dupoux</surname>, <given-names>E.</given-names></string-name> <year>2006</year>. <article-title>The acquisition of allophonic rules: Statistical learning with linguistic constraints</article-title>. <source>Cognition</source>, <volume>101</volume>(<issue>3</issue>), <fpage>B31</fpage>&#8211;<lpage>B41</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.cognition.2005.10.006</pub-id></mixed-citation></ref>
<ref id="B141"><label>141</label><mixed-citation publication-type="book"><string-name><surname>Pierrehumbert</surname>, <given-names>J.</given-names></string-name> <year>2001</year>. <chapter-title>Exemplar dynamics: Word frequency, lenition and contrast</chapter-title>. In: <string-name><surname>Bybee</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Hopper</surname>, <given-names>P.</given-names></string-name> (eds.), <source>Frequency and the emergence of linguistic structure</source>, <fpage>137</fpage>&#8211;<lpage>157</lpage>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1075/tsl.45.08pie</pub-id></mixed-citation></ref>
<ref id="B142"><label>142</label><mixed-citation publication-type="book"><string-name><surname>Pierrehumbert</surname>, <given-names>J.</given-names></string-name> <year>2002</year>. <chapter-title>Word-specific phonetics</chapter-title>. In: <string-name><surname>Gussenhoven</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Warner</surname>, <given-names>N.</given-names></string-name> (eds.), <source>Papers in laboratory phonology VII</source>, <fpage>101</fpage>&#8211;<lpage>140</lpage>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1515/9783110197105.101</pub-id></mixed-citation></ref>
<ref id="B143"><label>143</label><mixed-citation publication-type="journal"><string-name><surname>Pierrehumbert</surname>, <given-names>J.</given-names></string-name> <year>2016</year>. <article-title>Phonological representation: Beyond abstract versus episodic</article-title>. <source>Annual Review of Linguistics</source>, <volume>2</volume>, <fpage>33</fpage>&#8211;<lpage>52</lpage>. DOI: <pub-id pub-id-type="doi">10.1146/annurev-linguistics-030514-125050</pub-id></mixed-citation></ref>
<ref id="B144"><label>144</label><mixed-citation publication-type="book"><string-name><surname>Pinkerton</surname>, <given-names>S.</given-names></string-name> <year>1986</year>. <chapter-title>Quichean (Mayan) glottalized and nonglottalized stops: A phonetic study with implications for phonological universals</chapter-title>. In: <string-name><surname>Ohala</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Jaeger</surname>, <given-names>J.</given-names></string-name> (eds.), <source>Experimental phonology</source>, <fpage>125</fpage>&#8211;<lpage>139</lpage>. <publisher-loc>Orlando</publisher-loc>: <publisher-name>Academic Press</publisher-name>.</mixed-citation></ref>
<ref id="B145"><label>145</label><mixed-citation publication-type="journal"><string-name><surname>Pisoni</surname>, <given-names>D.</given-names></string-name> <year>1973</year>. <article-title>Auditory and phonetic memory codes in the discrimination of consonants and vowels</article-title>. <source>Perception &amp; Psychophysics</source>, <volume>13</volume>(<issue>2</issue>), <fpage>253</fpage>&#8211;<lpage>260</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03214136</pub-id></mixed-citation></ref>
<ref id="B146"><label>146</label><mixed-citation publication-type="journal"><string-name><surname>Pisoni</surname>, <given-names>D.</given-names></string-name> <year>1975</year>. <article-title>Auditory short-term memory and vowel perception</article-title>. <source>Memory &amp; Cognition</source>, <volume>3</volume>(<issue>1</issue>), <fpage>7</fpage>&#8211;<lpage>18</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03198202</pub-id></mixed-citation></ref>
<ref id="B147"><label>147</label><mixed-citation publication-type="journal"><string-name><surname>Pisoni</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Tash</surname>, <given-names>J.</given-names></string-name> <year>1974</year>. <article-title>Reaction times to comparisons within and across phonetic categories</article-title>. <source>Perception &amp; Psychophysics</source>, <volume>15</volume>(<issue>2</issue>), <fpage>285</fpage>&#8211;<lpage>290</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03213946</pub-id></mixed-citation></ref>
<ref id="B148"><label>148</label><mixed-citation publication-type="journal"><string-name><surname>Pitt</surname>, <given-names>M. A.</given-names></string-name>, &amp; <string-name><surname>Samuel</surname>, <given-names>A. G.</given-names></string-name> <year>1993</year>. <article-title>An empirical and meta-analytic evaluation of the phoneme identification task</article-title>. <source>Journal of Experimental Psychology: Human Perception and Performance</source>, <volume>19</volume>(<issue>4</issue>), <fpage>699</fpage>&#8211;<lpage>725</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0096-1523.19.4.699</pub-id></mixed-citation></ref>
<ref id="B149"><label>149</label><mixed-citation publication-type="journal"><string-name><surname>Port</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Leary</surname>, <given-names>A.</given-names></string-name> <year>2005</year>. <article-title>Against formal phonology</article-title>. <source>Language</source>, <volume>81</volume>(<issue>4</issue>), <fpage>927</fpage>&#8211;<lpage>964</lpage>. DOI: <pub-id pub-id-type="doi">10.1353/lan.2005.0195</pub-id></mixed-citation></ref>
<ref id="B150"><label>150</label><mixed-citation publication-type="journal"><string-name><surname>Quen&#233;</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>van Delft</surname>, <given-names>L.</given-names></string-name> <year>2010</year>. <article-title>Non-native durational patterns decrease speech intelligibility</article-title>. <source>Speech Communication</source>, <volume>52</volume>(<issue>11&#8211;12</issue>), <fpage>911</fpage>&#8211;<lpage>918</lpage>. (Non-native Speech Perception in Adverse Conditions). DOI: <pub-id pub-id-type="doi">10.1016/j.specom.2010.03.005</pub-id></mixed-citation></ref>
<ref id="B151"><label>151</label><mixed-citation publication-type="webpage"><collab>R Development Core Team</collab>. <year>2013</year>. <chapter-title>R: A language and environment for statistical computing [Computer software manual]</chapter-title>. <source>Computer program</source>. <publisher-name>Vienna, Austria</publisher-name>. (Version 3.0.1, retrieved from: <uri>http://www.R-project.org/</uri>).</mixed-citation></ref>
<ref id="B152"><label>152</label><mixed-citation publication-type="journal"><string-name><surname>Redford</surname>, <given-names>M. A.</given-names></string-name>, &amp; <string-name><surname>Diehl</surname>, <given-names>R. L.</given-names></string-name> <year>1999</year>. <article-title>The relative perceptual distinctiveness of initial and final consonants in CVC syllables</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>106</volume>, <fpage>1555</fpage>&#8211;<lpage>1565</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.427152</pub-id></mixed-citation></ref>
<ref id="B153"><label>153</label><mixed-citation publication-type="book"><string-name><surname>Renwick</surname>, <given-names>M.</given-names></string-name> <year>2014</year>. <source>The phonetics and phonology of contrast: The case of the Romanian vowel system</source>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Walter de Gruyter</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1515/9783110362770</pub-id></mixed-citation></ref>
<ref id="B154"><label>154</label><mixed-citation publication-type="journal"><string-name><surname>Repp</surname>, <given-names>B. H.</given-names></string-name>, &amp; <string-name><surname>Crowder</surname>, <given-names>R. G.</given-names></string-name> <year>1990</year>. <article-title>Stimulus order effects in vowel discrimination</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>88</volume>(<issue>5</issue>), <fpage>2080</fpage>&#8211;<lpage>2090</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.400105</pub-id></mixed-citation></ref>
<ref id="B155"><label>155</label><mixed-citation publication-type="book"><string-name><surname>Richards</surname>, <given-names>M.</given-names></string-name> <year>2003</year>. <source>Atlas ling&#252;&#237;stico de Guatemala</source>. <publisher-name>Instituto de Ling&#252;&#237;stico y Educaci&#243;n de la Universidad Rafael Land&#237;var</publisher-name>.</mixed-citation></ref>
<ref id="B156"><label>156</label><mixed-citation publication-type="journal"><string-name><surname>Rose</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>King</surname>, <given-names>L.</given-names></string-name> <year>2007</year>. <article-title>Speech error elicitation and co-occurrence restrictions in two Ethiopian Semitic languages</article-title>. <source>Language and Speech</source>, <volume>50</volume>(<issue>4</issue>), <fpage>451</fpage>&#8211;<lpage>504</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/00238309070500040101</pub-id></mixed-citation></ref>
<ref id="B157"><label>157</label><mixed-citation publication-type="thesis"><string-name><surname>Russell</surname>, <given-names>S.</given-names></string-name> <year>1997</year>. <source>Some acoustic characteristics of word initial pulmonic and glottalic stops in Mam</source> (Unpublished master&#8217;s thesis). <publisher-name>Simon Fraser University</publisher-name>.</mixed-citation></ref>
<ref id="B158"><label>158</label><mixed-citation publication-type="book"><string-name><surname>Sakoe</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>Chiba</surname>, <given-names>S.</given-names></string-name> <year>1971</year>. <chapter-title>A dynamic programming approach to continuous speech recognition</chapter-title>. In: <source>Proceedings of the seventh international congress on acoustics</source>, <volume>3</volume>, <fpage>65</fpage>&#8211;<lpage>69</lpage>. <publisher-loc>Budapest, Hungary</publisher-loc>: <publisher-name>Akademiai Kiado</publisher-name>.</mixed-citation></ref>
<ref id="B159"><label>159</label><mixed-citation publication-type="book"><string-name><surname>Sebasti&#225;n-Gall&#233;s</surname>, <given-names>N.</given-names></string-name> <year>2005</year>. <chapter-title>Cross-language speech perception</chapter-title>. In: <string-name><surname>Pisoni</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Remez</surname>, <given-names>R.</given-names></string-name> (eds.), <source>The handbook of speech perception</source>, <fpage>546</fpage>&#8211;<lpage>566</lpage>. <publisher-loc>Malden, MA</publisher-loc>: <publisher-name>Blackwell</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1002/9780470757024.ch22</pub-id></mixed-citation></ref>
<ref id="B160"><label>160</label><mixed-citation publication-type="journal"><string-name><surname>Shannon</surname>, <given-names>C. E.</given-names></string-name> <year>1948</year>, <month>July</month>. <article-title>A mathematical theory of communication</article-title>. <source>The Bell System Technical Journal</source>, <volume>27</volume>(<issue>3</issue>), <fpage>379</fpage>&#8211;<lpage>423</lpage>. DOI: <pub-id pub-id-type="doi">10.1002/j.1538-7305.1948.tb01338.x</pub-id></mixed-citation></ref>
<ref id="B161"><label>161</label><mixed-citation publication-type="book"><string-name><surname>Silverman</surname>, <given-names>D.</given-names></string-name> <year>2006</year>. <source>A critical introduction to phonology: Of sound, mind, and body</source>. <publisher-loc>London &amp; New York</publisher-loc>: <publisher-name>Continuum</publisher-name>.</mixed-citation></ref>
<ref id="B162"><label>162</label><mixed-citation publication-type="book"><string-name><surname>Silverman</surname>, <given-names>D.</given-names></string-name> <year>2012</year>. <source>Neutralization</source>. <publisher-loc>Cambridge, UK</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1017/CBO9781139013895</pub-id></mixed-citation></ref>
<ref id="B163"><label>163</label><mixed-citation publication-type="journal"><string-name><surname>Smits</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Sereno</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Jongman</surname>, <given-names>A.</given-names></string-name> <year>2006</year>. <article-title>Categorization of sounds</article-title>. <source>Journal of Experimental Psychology: Human Perception and Performance</source>, <volume>32</volume>(<issue>3</issue>), <fpage>733</fpage>&#8211;<lpage>754</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/e501882009-357</pub-id></mixed-citation></ref>
<ref id="B164"><label>164</label><mixed-citation publication-type="book"><string-name><surname>Steriade</surname>, <given-names>D.</given-names></string-name> <year>2001</year>. <chapter-title>Directional asymmetries in place assimilation: A perceptual account</chapter-title>. In: <string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Hume</surname>, <given-names>E.</given-names></string-name> (eds.), <source>The role of speech perception in phonology</source>, <fpage>219</fpage>&#8211;<lpage>250</lpage>. <publisher-loc>New York</publisher-loc>: <publisher-name>Academic Press</publisher-name>.</mixed-citation></ref>
<ref id="B165"><label>165</label><mixed-citation publication-type="book"><string-name><surname>Steriade</surname>, <given-names>D.</given-names></string-name> <year>2009</year>. <chapter-title>The phonology of perceptibility effects: The P-map and its consequences for constraint organization</chapter-title>. In: <string-name><surname>Hanson</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Inkelas</surname>, <given-names>S.</given-names></string-name> (eds.), <source>The nature of the word: Studies in honor of Paul Kiparsky</source>, <fpage>151</fpage>&#8211;<lpage>179</lpage>. <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.7551/mitpress/9780262083799.003.0007</pub-id></mixed-citation></ref>
<ref id="B166"><label>166</label><mixed-citation publication-type="journal"><string-name><surname>Stevenson</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Zamuner</surname>, <given-names>T.</given-names></string-name> <year>2017</year>. <article-title>Gradient phonological relationships: Evidence from vowels in French</article-title>. <source>Glossa</source>, <volume>2</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>22</lpage>. DOI: <pub-id pub-id-type="doi">10.5334/gjgl.162</pub-id></mixed-citation></ref>
<ref id="B167"><label>167</label><mixed-citation publication-type="book"><string-name><surname>Surendran</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Niyogi</surname>, <given-names>P.</given-names></string-name> <year>2003</year>. <source>Measuring the usefulness (functional load) of phonological contrasts</source> (Tech. Rep.). <publisher-loc>Chicago</publisher-loc>: <publisher-name>Department of Computer Science, University of Chicago</publisher-name>. (Technical Report TR-2003).</mixed-citation></ref>
<ref id="B168"><label>168</label><mixed-citation publication-type="book"><string-name><surname>Surendran</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Niyogi</surname>, <given-names>P.</given-names></string-name> <year>2006</year>. <chapter-title>Quantifying the functional load of phonemic oppositions, distinctive features, and suprasegmentals</chapter-title>. In: <string-name><surname>Thomsen</surname>, <given-names>O. N.</given-names></string-name> (ed.), <source>Competing models of linguistic change: Evolution and beyond</source>, <fpage>43</fpage>&#8211;<lpage>58</lpage>. <publisher-loc>Amsterdam</publisher-loc>: <publisher-name>John Benjamins</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1075/cilt.279.05sur</pub-id></mixed-citation></ref>
<ref id="B169"><label>169</label><mixed-citation publication-type="journal"><string-name><surname>Sze</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Rickard Liow</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Yap</surname>, <given-names>M.</given-names></string-name> <year>2014</year>. <article-title>The Chinese Lexicon Project: A repository of lexical decision behavioral responses for 2,500 Chinese characters</article-title>. <source>Behavior Research Methods</source>, <volume>46</volume>(<issue>1</issue>), <fpage>263</fpage>&#8211;<lpage>273</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/s13428-013-0355-9</pub-id></mixed-citation></ref>
<ref id="B170"><label>170</label><mixed-citation publication-type="thesis"><string-name><surname>Tang</surname>, <given-names>K.</given-names></string-name> <year>2015</year>. <source>Naturalistic speech misperception</source> (Unpublished doctoral dissertation). <publisher-name>University College London</publisher-name>.</mixed-citation></ref>
<ref id="B171"><label>171</label><mixed-citation publication-type="confproc"><string-name><surname>Tang</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Ajsivinac Sian</surname>, <given-names>J.</given-names></string-name> <year>2015</year>, <conf-date>July 13&#8211;17</conf-date>. <article-title>Modelling phonetic and phonological variation with &#8216;small&#8217; data: Evidence from Kaqchikel Mayan</article-title>. <conf-name>Presentation at Laboratory Phonology 15 conference</conf-name>.</mixed-citation></ref>
<ref id="B172"><label>172</label><mixed-citation publication-type="journal"><string-name><surname>Tang</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Bennett</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Ajsivinac Sian</surname>, <given-names>J.</given-names></string-name> (in preparation). <article-title>Contextual predictability influences word duration in a morphologically complex language (Kaqchikel Mayan)</article-title>.</mixed-citation></ref>
<ref id="B173"><label>173</label><mixed-citation publication-type="book"><string-name><surname>Tang</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Nevins</surname>, <given-names>A.</given-names></string-name> <year>2014</year>. <chapter-title>Measuring segmental and lexical trends in a corpus of naturalistic speech</chapter-title>. In: <string-name><surname>Huang</surname>, <given-names>H.-L.</given-names></string-name>, <string-name><surname>Poole</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Rysling</surname>, <given-names>A.</given-names></string-name> (eds.), <volume>2</volume>, <fpage>153</fpage>&#8211;<lpage>166</lpage>. <publisher-loc>Amherst, MA</publisher-loc>: <publisher-name>GLSA</publisher-name>.</mixed-citation></ref>
<ref id="B174"><label>174</label><mixed-citation publication-type="journal"><string-name><surname>Tang</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Nevins</surname>, <given-names>A.</given-names></string-name> (in preparation). <article-title>A graceful degradation account of lexical retrieval: Evidence from naturalistic misperception</article-title>.</mixed-citation></ref>
<ref id="B175"><label>175</label><mixed-citation publication-type="journal"><string-name><surname>Tilsen</surname>, <given-names>S.</given-names></string-name> <year>2016</year>. <article-title>Selection and coordination: The articulatory basis for the emergence of phonological structure</article-title>. <source>Journal of Phonetics</source>, <volume>55</volume>, <fpage>53</fpage>&#8211;<lpage>77</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2015.11.005</pub-id></mixed-citation></ref>
<ref id="B176"><label>176</label><mixed-citation publication-type="book"><string-name><surname>Trubetzkoy</surname>, <given-names>N.</given-names></string-name> <year>1939</year>. <source>Grundz&#252;ge der Phonologie</source>. <chapter-title>Travaux du cercle linguistique de Prague. (English translation published 1969 as <italic>Principles of phonology</italic></chapter-title>, trans. <string-name><given-names>C.A.M.</given-names> <surname>Baltaxe</surname></string-name>. <publisher-loc>Berkeley</publisher-loc>: <publisher-name>University of California Press</publisher-name>).</mixed-citation></ref>
<ref id="B177"><label>177</label><mixed-citation publication-type="journal"><string-name><surname>van Heuven</surname>, <given-names>W. J. B.</given-names></string-name>, <string-name><surname>Mandera</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Keuleers</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Brysbaert</surname>, <given-names>M.</given-names></string-name> <year>2014</year>. <article-title>SUBTLEX-UK: A new and improved word frequency database for British English</article-title>. <source>The Quarterly Journal of Experimental Psychology</source>, <volume>67</volume>(<issue>6</issue>), <fpage>1176</fpage>&#8211;<lpage>1190</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/17470218.2013.850521</pub-id></mixed-citation></ref>
<ref id="B178"><label>178</label><mixed-citation publication-type="journal"><string-name><surname>Vitevitch</surname>, <given-names>M.</given-names></string-name> <year>2002</year>. <article-title>Naturalistic and experimental analyses of word frequency and neighborhood density effects in slips of the ear</article-title>. <source>Language and speech</source>, <volume>45</volume>(<issue>4</issue>), <fpage>407</fpage>&#8211;<lpage>434</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/00238309020450040501</pub-id></mixed-citation></ref>
<ref id="B179"><label>179</label><mixed-citation publication-type="journal"><string-name><surname>Vitevitch</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Luce</surname>, <given-names>P. A.</given-names></string-name> <year>2016</year>. <article-title>Phonological neighborhood effects in spoken word perception and production</article-title>. <source>Annual Review of Linguistics</source>, <volume>2</volume>, <fpage>75</fpage>&#8211;<lpage>94</lpage>. DOI: <pub-id pub-id-type="doi">10.1146/annurev-linguistics-030514-124832</pub-id></mixed-citation></ref>
<ref id="B180"><label>180</label><mixed-citation publication-type="journal"><string-name><surname>Wang</surname>, <given-names>M. D.</given-names></string-name>, &amp; <string-name><surname>Bilger</surname>, <given-names>R. C.</given-names></string-name> <year>1973</year>. <article-title>Consonant confusions in noise: A study of perceptual features</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>54</volume>(<issue>5</issue>), <fpage>1248</fpage>&#8211;<lpage>1266</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1914417</pub-id></mixed-citation></ref>
<ref id="B181"><label>181</label><mixed-citation publication-type="thesis"><string-name><surname>Wedel</surname>, <given-names>A.</given-names></string-name> <year>2004</year>. <source>Self-organization and categorical behavior in phonology</source> (Unpublished doctoral dissertation). <publisher-name>UC Santa Cruz</publisher-name>.</mixed-citation></ref>
<ref id="B182"><label>182</label><mixed-citation publication-type="journal"><string-name><surname>Wedel</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Jackson</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Kaplan</surname>, <given-names>A.</given-names></string-name> <year>2013</year>. <article-title>Functional load and the lexicon: Evidence that syntactic category and frequency relationships in minimal lemma pairs predict the loss of phoneme contrasts in language change</article-title>. <source>Language and speech</source>, <volume>56</volume>(<issue>3</issue>), <fpage>395</fpage>&#8211;<lpage>417</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0023830913489096</pub-id></mixed-citation></ref>
<ref id="B183"><label>183</label><mixed-citation publication-type="journal"><string-name><surname>Wedel</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Kaplan</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Jackson</surname>, <given-names>S.</given-names></string-name> <year>2013</year>. <article-title>High functional load inhibits phonological contrast loss: A corpus study</article-title>. <source>Cognition</source>, <volume>128</volume>(<issue>2</issue>), <fpage>179</fpage>&#8211;<lpage>186</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.cognition.2013.03.002</pub-id></mixed-citation></ref>
<ref id="B184"><label>184</label><mixed-citation publication-type="journal"><string-name><surname>Werker</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Logan</surname>, <given-names>J.</given-names></string-name> <year>1985</year>. <article-title>Cross-language evidence for three factors in speech perception</article-title>. <source>Perception &amp; Psychophysics</source>, <volume>37</volume>(<issue>1</issue>), <fpage>35</fpage>&#8211;<lpage>44</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03207136</pub-id></mixed-citation></ref>
<ref id="B185"><label>185</label><mixed-citation publication-type="journal"><string-name><surname>Werker</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Tees</surname>, <given-names>R.</given-names></string-name> <year>1984a</year>. <article-title>Cross-language speech perception: Evidence for perceptual reorganization during the first year of life</article-title>. <source>Infant behavior and development</source>, <volume>7</volume>(<issue>1</issue>), <fpage>49</fpage>&#8211;<lpage>63</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/S0163-6383(02)00093-0</pub-id></mixed-citation></ref>
<ref id="B186"><label>186</label><mixed-citation publication-type="journal"><string-name><surname>Werker</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Tees</surname>, <given-names>R.</given-names></string-name> <year>1984b</year>. <article-title>Phonemic and phonetic factors in adult cross-language speech perception</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>75</volume>(<issue>6</issue>), <fpage>1866</fpage>&#8211;<lpage>1878</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.390988</pub-id></mixed-citation></ref>
<ref id="B187"><label>187</label><mixed-citation publication-type="journal"><string-name><surname>Whalen</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>McDonough</surname>, <given-names>J.</given-names></string-name> <year>2015</year>. <article-title>Taking the laboratory into the field</article-title>. <source>Annual Review of Linguistics</source>, <volume>1</volume>(<issue>1</issue>), <fpage>395</fpage>&#8211;<lpage>415</lpage>. DOI: <pub-id pub-id-type="doi">10.1146/annurev-linguist-030514-124915</pub-id></mixed-citation></ref>
<ref id="B188"><label>188</label><mixed-citation publication-type="journal"><string-name><surname>Wright</surname>, <given-names>C. E.</given-names></string-name> <year>1979</year>. <article-title>Duration differences between rare and common words and their implications for the interpretation of word frequency effects</article-title>. <source>Memory &amp; Cognition</source>, <volume>7</volume>(<issue>6</issue>), <fpage>411</fpage>&#8211;<lpage>419</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03198257</pub-id></mixed-citation></ref>
<ref id="B189"><label>189</label><mixed-citation publication-type="book"><string-name><surname>Wright</surname>, <given-names>R.</given-names></string-name> <year>2004</year>. <chapter-title>A review of perceptual cues and cue robustness</chapter-title>. In: <string-name><surname>Hayes</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Kirchner</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Steriade</surname>, <given-names>D.</given-names></string-name> (eds.), <source>Phonetically based phonology</source>, <fpage>34</fpage>&#8211;<lpage>57</lpage>. <publisher-loc>Cambridge, UK</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1017/CBO9780511486401.002</pub-id></mixed-citation></ref>
<ref id="B190"><label>190</label><mixed-citation publication-type="journal"><string-name><surname>Wright</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Hargus</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Davis</surname>, <given-names>K.</given-names></string-name> <year>2002</year>. <article-title>On the categorization of ejectives: Data from Witsuwit&#8217;en</article-title>. <source>Journal of the International Phonetic Association</source>, <volume>32</volume>(<issue>1</issue>), <fpage>43</fpage>&#8211;<lpage>77</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0025100302000142</pub-id></mixed-citation></ref>
<ref id="B191"><label>191</label><mixed-citation publication-type="journal"><string-name><surname>Xu</surname>, <given-names>Y.</given-names></string-name> <year>2010</year>. <article-title>In defense of lab speech</article-title>. <source>Journal of Phonetics</source>, <volume>38</volume>(<issue>3</issue>), <fpage>329</fpage>&#8211;<lpage>336</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2010.04.003</pub-id></mixed-citation></ref>
<ref id="B192"><label>192</label><mixed-citation publication-type="thesis"><string-name><surname>Yao</surname>, <given-names>Y.</given-names></string-name> <year>2011</year>. <source>The effects of phonological neighborhoods on pronunciation variation in conversational speech</source> (Unpublished doctoral dissertation). <publisher-name>University of California</publisher-name>, <publisher-loc>Berkeley</publisher-loc>.</mixed-citation></ref>
<ref id="B193"><label>193</label><mixed-citation publication-type="journal"><string-name><surname>Yap</surname>, <given-names>M. J.</given-names></string-name>, <string-name><surname>Sibley</surname>, <given-names>D. E.</given-names></string-name>, <string-name><surname>Balota</surname>, <given-names>D. A.</given-names></string-name>, <string-name><surname>Ratcliff</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Rueckl</surname>, <given-names>J.</given-names></string-name> <year>2015</year>. <article-title>Responding to nonwords in the lexical decision task: Insights from the English Lexicon Project</article-title>. <source>Journal of Experimental Psychology: Learning, Memory, and Cognition</source>, <volume>41</volume>(<issue>3</issue>), <fpage>597</fpage>&#8211;<lpage>613</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/xlm0000064</pub-id></mixed-citation></ref>
<ref id="B194"><label>194</label><mixed-citation publication-type="journal"><string-name><surname>Yu</surname>, <given-names>A. C. L.</given-names></string-name> <year>2011</year>. <article-title>On measuring phonetic precursor robustness: A response to Moreton</article-title>. <source>Phonology</source>, <volume>28</volume>(<issue>3</issue>), <fpage>491</fpage>&#8211;<lpage>518</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0952675711000236</pub-id></mixed-citation></ref>
<ref id="B195"><label>195</label><mixed-citation publication-type="book"><string-name><surname>Zipf</surname>, <given-names>G. K.</given-names></string-name> <year>1935</year>. <source>The psycho-biology of language</source>. <publisher-loc>Boston</publisher-loc>: <publisher-name>Houghton Mifflin</publisher-name>.</mixed-citation></ref>
</ref-list>
</back>
</article>