<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.1" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology: Journal of the Association for Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/labphon.240</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>The perceptual filtering of predictable coarticulation in exemplar memory</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Manker</surname>
<given-names>Jonathan</given-names>
</name>
<email>jonathan.manker@rice.edu</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>Department of Linguistics, Rice University, Houston, TX, US</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2020-11-19">
<day>19</day>
<month>11</month>
<year>2020</year>
</pub-date>
<pub-date pub-type="collection">
<year>2020</year>
</pub-date>
<volume>11</volume>
<issue>1</issue>
<elocation-id>20</elocation-id>
<history>
<date date-type="received" iso-8601-date="2019-10-18">
<day>18</day>
<month>10</month>
<year>2019</year>
</date>
<date date-type="accepted" iso-8601-date="2020-09-28">
<day>28</day>
<month>09</month>
<year>2020</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2020 The Author(s)</copyright-statement>
<copyright-year>2020</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.5334/labphon.240/"/>
<abstract>
<p>Exemplar models of word representations have remained ambivalent or impressionistic as to precisely what veridical auditory information is stored in individual word exemplars. Earlier models (<xref ref-type="bibr" rid="B17">Johnson, 1997b</xref>) suggest all perceived information was stored in memory, whereas more recent proposals (<xref ref-type="bibr" rid="B36">Pierrehumbert, 2002</xref>; <xref ref-type="bibr" rid="B5">Goldinger, 2007</xref>) suggest some degree of abstraction occurs in storing particular exemplars. Findings from the phonetic accommodation paradigm (<xref ref-type="bibr" rid="B6">Goldinger, 1998</xref>; <xref ref-type="bibr" rid="B32">Nielsen, 2011</xref>, etc.) suggest that the accumulation of new exemplars may drive the spread of sound change. At the same time, some theories of sound change suggest that perceptual biases serve as a starting point for change (<xref ref-type="bibr" rid="B34">Ohala, 1981</xref>, <xref ref-type="bibr" rid="B35">1983</xref>). The current study investigates how perceptual biases, such as the predictability of coarticulation, can shape the contents of exemplars. The experimental results suggest that an <italic>expected</italic> phonetic alteration, such as f0 raising on vowels following voiceless consonants&#8212;a predictable coarticulatory effect&#8212;is more likely to undergo some degree of abstraction when stored in exemplar memory, whereas <italic>unexpected</italic> phonetic detail (e.g., f0 raising following voiced consonants) is more faithfully stored or maintained for longer in memory. These findings suggest perceptual biases that could shape pools of exemplars, leading to different expectations for conditioned versus unconditioned sound changes.</p>
</abstract>
<kwd-group>
<kwd>Exemplar theory</kwd>
<kwd>discrimination</kwd>
<kwd>speech perception</kwd>
<kwd>coarticulation</kwd>
<kwd>sound change</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>1. Background: Exemplars</title>
<sec>
<title>1.1. Exemplars store fine phonetic details</title>
<p>Phonologists have long attempted to describe and model how words are represented in the mental lexicon. One of the most enduring proposals is that words are stored as abstract, formal representations, composed of strings of contrastive segments, a principle borrowed from generative phonology (<xref ref-type="bibr" rid="B2">Chomsky &amp; Halle, 1968</xref>) and applied to models of speech perception involving normalization (<xref ref-type="bibr" rid="B3">Gerstman, 1968</xref>; <xref ref-type="bibr" rid="B45">Tranm&#252;ller, 1981</xref>, etc.). Contemporaneous to these developments, phoneticians were becoming aware of the vast acoustic variability of the speech signal. Liberman et al. (<xref ref-type="bibr" rid="B26">1967</xref>) demonstrated the variability of particular phonemes produced by speakers due to coarticulatory effects, while Stevens (<xref ref-type="bibr" rid="B41">1972</xref>) and Klatt (<xref ref-type="bibr" rid="B23">1979</xref>) noted the variability due to idiosyncrasies of a speaker&#8217;s voice, gender, and vocal tract size (even among speakers from the same speech communities). This acoustic variability proved problematic without an additional system by which phonological representations of words could be extracted from the messy acoustic signal. The solution was in the normalization of the speech signal: Coarticulatory effects were undone, and the idiosyncratic features of talkers&#8217; voices were stripped away by the listener, allowing the activation of an invariant phonological form.</p>
<p>In the following decades, exemplar models challenged the concept of abstract phonological forms as the mental representations of words. The core principle of exemplar-based models is that percepts&#8212;in the case of speech perception, words&#8212;are stored in memory as collections of individual labeled instances, rather than as a single abstraction. Goldinger (<xref ref-type="bibr" rid="B5">1996</xref>) provided evidence that exemplars were used in speech perception, finding that subjects were more accurate in identifying whether a word had been repeated or not if the repetitions were produced in the same voice. This suggests not only that voice information was stored in memory, but that this information was an integrated part of the percept along with the phonological form of the word itself. If only an abstract representation had been activated, with voice information stripped away, no voice effect should have been found.</p>
<p>These findings led to a line of research in <italic>phonetic accommodation</italic> which extended and further corroborated the predictions of the exemplar model. The theory of phonetic accommodation asserts that an individual&#8217;s pronunciation will drift towards that exhibited by his or her interlocutors after being exposed to it. Phonetic accommodation is predicted by exemplar theory because if prior stored instances of words are what comprise an individual&#8217;s mental representation of a word, a new instance will then (slightly) shift the mean pronunciation of future utterances of that word. Goldinger (<xref ref-type="bibr" rid="B6">1998</xref>) observed just such an effect in an AXB task, such that speakers&#8217; repetitions of words (B) heard after the stimulus (X) sounded more like the stimulus than the original pronunciation (A). Furthermore, Goldinger found other effects that suggest the structure of exemplar clouds. Immediate shadowing of the stimuli produced a stronger imitative effect than delayed shadowing, suggesting new exemplars fade from memory with time. Secondly, more repetitions of words resulted in stronger accommodation, suggesting a greater number of new exemplars have a greater ability to shift the previous production averages. Lastly, low frequency words also displayed greater accommodation, suggesting a smaller pre-existing cloud of exemplars would be more prone to change than those containing more prior exemplars. While Goldinger&#8217;s (<xref ref-type="bibr" rid="B6">1998</xref>) accommodation study judged similarity to the model based on the qualitative impressions of a panel of judges, later studies were able to quantify phonetic drift due to accommodation, measuring small changes in particular phonetic features such as VOT (<xref ref-type="bibr" rid="B39">Shockley, Sabadini, &amp; Fowler, 2004</xref>; <xref ref-type="bibr" rid="B32">Nielsen, 2011</xref>) and vowel quality (<xref ref-type="bibr" rid="B44">Tilsen, 2009</xref>).</p>
<p>While many studies have established strong evidence for the existence of word exemplars and their relevance in speech perception, it is less clear precisely what information is stored in one. Generally speaking, a standard exemplar model assumes that exemplars are fairly veridical representations, being stored &#8220;as they occur, without any abstraction at all&#8221; (<xref ref-type="bibr" rid="B18">Johnson, 2007, p. 27</xref>). However, the abstraction that Johnson mentions refers to the transformation of newly perceived percepts into prototypical forms, as needed for speech recognition, a separate phenomenon from exemplars themselves undergoing some sort of perceptual transformation or abstraction while stored in memory as part of clouds of instances of particular words. Nevertheless, studies by Johnson (<xref ref-type="bibr" rid="B16">1997a</xref>) and (<xref ref-type="bibr" rid="B17">1997b</xref>) propose exemplars with purely veridical information, though with varying degrees. Johnson (<xref ref-type="bibr" rid="B16">1997a</xref>) considers reduced representations of vowel exemplars composed only of formant values. However, Johnson (<xref ref-type="bibr" rid="B17">1997b</xref>) argues for the storage of &#8216;auditory spectra&#8217; containing all auditory information that the ear gleans from the acoustic signal. He states that this more detailed representation &#8220;is realistic because it is based on psychoacoustic data, and it also avoids making assumptions about which of the many potential acoustic features should be measured and kept in an exemplar of heard speech&#8221; (2007, p. 34), though he concedes that more data-driven evidence could lend support for a more compact representation.</p>
<p>Others propose models which simultaneously contain veridical exemplars and abstract phonological representations. For example, Pierrehumbert (<xref ref-type="bibr" rid="B36">2002</xref>, <xref ref-type="bibr" rid="B37">2016</xref>) argues in favor of a hybrid model containing both phonological encoding and exemplar representations. This helps to account both for generalizations such as Neogrammarian type sound changes which affect all instances of a particular phoneme, as well as word-specific phonetics that might be influenced by word frequency, a problem better handled by exemplar representations. McLennan and Luce (<xref ref-type="bibr" rid="B31">2005</xref>) and Luce and McLennan (<xref ref-type="bibr" rid="B27">2005</xref>) find that both abstract representations and exemplars may co-exist, being activated at different stages in speech perception. While such hybrid models include both phonological and exemplar levels of representation, it is unclear if exemplars in these models would have purely veridical information or if they may include some degree of abstraction as well.</p>
<p>Goldinger (<xref ref-type="bibr" rid="B5">1996</xref>) does not directly define the contents of exemplars, although he describes the perceptual process of encoding voice information in episodes as being automatic, which could suggest there is no specificity in what details are stored. Nevertheless, he also notes that episodic encoding of phonetic details may occur &#8220;only to the extent that they matter in original processing,&#8221; and that these memory traces will usually &#8220;emphasize elements of meaning, not perception&#8221; (p. 1180). Goldinger (<xref ref-type="bibr" rid="B7">2007</xref>) further considers the question of what exemplars encode, asserting that any &#8216;raw data&#8217; will necessarily undergo some abstract transformation, such that &#8220;each stored &#8216;exemplar&#8217; is actually a product of perceptual input combined with prior knowledge, the precise balance likely affected by many factors&#8221; (p. 50). Hawkins (<xref ref-type="bibr" rid="B10">2003</xref>, <xref ref-type="bibr" rid="B11">2010</xref>), proposes that exemplars with fine phonetic detail are first stored in memory, but are processed for signal-to-structure mapping, only to the extent needed to achieve an understanding of the linguistic meaning. This can be achieved a number of ways depending on the context, whereas processing of the individual words, phonemes, or subphonemic details may not be necessary depending on the context. In some cases, top-down processing will be faster in identifying words and phonemes as opposed to actual processing of the speech signal, in which case the veridical phonetic details may not have been used to understand the linguistic meaning. This suggests a possible bias in the storage and maintenance of different phonetic information present in the speech signal which might in some way be modulated by contextual information.</p>
<p>Some previous work has similarly indicated that the availability of top-down information could be a factor in the storage and maintenance of phonetic detail. Nye and Fowler (<xref ref-type="bibr" rid="B33">2003</xref>) is a rare case of an accommodation study which used sentence rather than word stimuli, and as such, provided insight into how top-down information influences phonetic accommodation. In their experiment, subjects were presented with nonce word stimuli that varied in how closely they approximated English, particularly with regard to phonotactic similarity&#8212;e.g., low orders of approximation grossly violated English phonotactic structure, such as [&#601; tb o&#618;m&#603;k&#652;nd pr&#601;n v&#650;&#643;&#601;l], whereas higher order stimuli closely resembled English, such as [hiz &#601; p&#618;nto &#230;nd hi fot&#601;gr&#230;s w&#652;nd&#602;f&#601;lli]. The results showed that subjects more closely imitated the <italic>lower</italic> orders of approximation, which suggested that more information about the higher order stimuli may have undergone some degree of abstraction due to high level linguistic knowledge of English, allowing for less detailed processing of higher order stimuli.</p>
<p>Manker (<xref ref-type="bibr" rid="B29">2019</xref>) directly assessed the effect of contextual knowledge on the storage of phonetic details in exemplar memory. In a discrimination task, subjects heard sentences with one of the words repeated, and were asked to determine if the word sounded exactly the same or somewhat different when repeated. The results showed that subjects displayed better discrimination when the words had been heard in an <italic>unpredictable</italic> context, e.g., &#8220;Joe turned and saw the <italic>cabins</italic>,&#8221; as opposed to a <italic>predictable</italic> context, e.g., &#8220;Pioneers built log <italic>cabins</italic>.&#8221; In a second accommodation experiment with the same stimuli, it was revealed that subjects also displayed better VOT and pitch contour accommodation for words that had been heard in unpredictable context, corroborating the results of the discrimination task. Ultimately, this suggested that more detailed, veridical exemplars were stored or maintained in memory for unpredictable words, because more reliance on processing of fine phonetic detail was needed in the first place in order to identify these words, whereas the recognition of predictable words could be achieved via top-down processing using contextual information.</p>
<p>Thus, these studies showed that semantic contextual information is one possible bias affecting what details might be stored in exemplar memory. Other linguistic information may also make certain details of the speech signal prone to being ignored or abstracted. The current study considers the effect of coarticulation in a similar way: If a particular coarticulatory effect is expected, does it need to be processed, stored, and maintained in exemplar memory when such detail could easily be abstracted or reconstructed based on higher level linguistic knowledge?</p>
<p>To answer this question, I conducted a discrimination task that compared listeners&#8217; abilities to store and maintain the acoustic details of expected coarticulation compared to unexpected acoustic detail. This should test whether coarticulatory predictability modulates the details that are stored in exemplar memory, by either stripping them away or abstracting them. I specifically examined the coarticulatory phenomenon of f0 raising versus lowering following voiceless and voiced consonants respectively, a phenomenon which can lead to tonogenesis following neutralization of consonant voicing (<xref ref-type="bibr" rid="B14">Hombert, Ohala, &amp; Ewan, 1979</xref>). I hypothesized that subjects would show better discrimination of stimuli that differ in acoustic detail that would not arise from coarticulation, which would suggest more veridical details of the auditory signal are stored in exemplar memory for unpredictable speech.</p>
</sec>
</sec>
<sec sec-type="methods">
<title>2. Methodology</title>
<sec>
<title>2.1. Voicing and F0</title>
<p>Phoneticians have established that a consonant&#8217;s voicing can slightly perturb the f0 of the following vowel (<xref ref-type="bibr" rid="B15">House &amp; Fairbanks, 1953</xref>; <xref ref-type="bibr" rid="B25">Lehiste &amp; Peterson, 1961</xref>; <xref ref-type="bibr" rid="B13">Hombert &amp; Ladefoged, 1976</xref>; <xref ref-type="bibr" rid="B14">Hombert et al., 1979</xref>; <xref ref-type="bibr" rid="B19">Kingston, 1989</xref>; <xref ref-type="bibr" rid="B21">Kingston &amp; Diehl, 1994</xref>; <xref ref-type="bibr" rid="B22">Kingston, Diehl, Kirk, &amp; Castleman, 2008</xref>; <xref ref-type="bibr" rid="B20">Kingston, 2011</xref>; <xref ref-type="bibr" rid="B38">Ratliff, 2015</xref>). Following a voiceless consonant, the f0 of the vowel tends to be raised 5&#8211;10 Hz briefly before sloping down to its intended target, whereas voiced consonants cause a similar pattern of pitch lowering (Figure <xref ref-type="fig" rid="F1">1</xref>). According to Hombert et al. (<xref ref-type="bibr" rid="B14">1979</xref>), this happens either due to the increased oral pressure which results from voicing, leading to lower pressure in the vocal folds and thus lower pitch, or due to the vocal cord tension needed for consonant voicing which affects the tension of the vocal folds during the following vowel as well. Ultimately, this phenomenon often sows the seeds of tonogenesis, where an older voicing contrast, e.g., /pa/ ~ /ba/, is enhanced with pitch differences, thus /p&#225;/ versus /b&#224;/, and is then lost with the survival of the new tone contrast, /p&#225;/ ~ /p&#224;/ (<xref ref-type="bibr" rid="B14">Hombert et al., 1979</xref>). Such a development, whether completed or not, has been observed in a diverse body of languages, including several Mon Khmer languages (<xref ref-type="bibr" rid="B42">Svantesson, 1991</xref>), Chadic languages (<xref ref-type="bibr" rid="B46">Wolff, 1987</xref>), Punjabi (<xref ref-type="bibr" rid="B4">Gill &amp; Gleason, 1972</xref>), Cham (<xref ref-type="bibr" rid="B43">Thurgood, 1999</xref>), and many others.</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Effect of voicing and voicelessness on f0, reproduced from Hombert et al. (<xref ref-type="bibr" rid="B14">1979</xref>).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75830/"/>
</fig>
<p>Given these observations, in the current experiment, I investigated whether listeners more successfully discriminated a word and its repetition with an f0 contour that would not arise from coarticulation as opposed to a word and repetition with a coarticularily expected f0 contour. That is to say, listeners were expected to store the f0 details of a syllable /ba&#7621;/ (indicating a slightly raised f0 at the beginning) better than /pa&#7621;/, where such a pitch contour is expected (Figure <xref ref-type="fig" rid="F2">2</xref>). The raised f0 contour will be the predictable result of coarticulation from the preceding voiceless stop, and such details may then not survive in exemplar storage.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Demonstration of hypothesized predictability-based perceptual bias.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75831/"/>
</fig>
</sec>
<sec>
<title>2.2. Participants</title>
<p>One hundred subjects in two counterbalanced groups of 50 were recruited via Amazon Mechanical Turk, an online platform also used for conducting the experiment (<xref ref-type="bibr" rid="B49">Yu &amp; Lee, 2014</xref> find similar results in speech perception experiments run in person versus using Amazon Mechanical Turk). Subjects were required to be native speakers of English located in the United States and without speech or hearing disorders. They were compensated with $3.50 for the approximately 20-minute experiment.</p>
</sec>
<sec>
<title>2.3. Procedure</title>
<p>The experiment consisted of 200 trials for each subject with two basic types of stimuli. Ninety-six of the stimuli (8 examples &#215; 6 different phonemes &#215; 2 repetition conditions, &#8216;same&#8217; or &#8216;different&#8217;) were AX discrimination trials in which the subject heard a word spoken in isolation, followed by a one-second pause, two seconds of multi-speaker babble (as a distractor), and then a repetition of the initial word (Figure <xref ref-type="fig" rid="F3">3</xref>). When repeated, the word had either the exact same f0 contour as the initial hearing, or the beginning was raised slightly mimicking the effect of pitch raising following a voiceless consonant (see Section 2.3. for more details). Then, the subject was asked whether the repetition sounded exactly the &#8216;same&#8217; or &#8216;different&#8217; than the initial hearing.</p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Demonstration of both trial types.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75832/"/>
</fig>
<p>Of these 96 discrimination stimuli, there were 48 which were in fact different when repeated, and 48 which were the same. Each of these 48 then was evenly divided into 24 stimuli with target words beginning with voiceless stops and 24 stimuli with voiced stops. These included eight tokens for each of six different phonemes at three different places of articulation, /p t k b d g/.</p>
<p>The other 104 stimuli, which thus occurred slightly more than 50% of the time, were word identification trials. After hearing the target word, there was a pause and subjects were asked to type the word they heard. It is important to note that subjects would not know which type of stimulus they were presented with until after hearing the initial word, at which time they either heard a repetition or not. These trials were intended to be distractors, to encourage the subjects to listen to speech more &#8216;naturally&#8217;&#8212;that is, to identify words rather than focusing on subphonemic details. For example, there could be some concern that the subjects might become aware that &#8216;different&#8217; always means a raised pitch contour, and they would start listening at a more acoustic level, possibly overriding or weakening any higher-level perceptual bias that might arise in a natural listening setting. However, this is mostly a concern of achieving a false negative, as such a phenomenon would result in no difference between how the voiced and voiceless stimuli are perceived. Any significant difference then between the perception of the voiced and voiceless stimuli would indicate an effect of the initial consonant&#8217;s voicing. A similar approach was used in Manker (<xref ref-type="bibr" rid="B29">2019</xref>), though it is unclear whether contextual level effects would not occur without the distractors. The procedure involved in the two types of stimuli are represented visually in Figure <xref ref-type="fig" rid="F3">3</xref>, while a breakdown of the different stimuli conditions is shown in the tree in Figure <xref ref-type="fig" rid="F4">4</xref>.</p>
<fig id="F4">
<label>Figure 4</label>
<caption>
<p>Tree showing the breakup of different stimuli.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75833/"/>
</fig>
<p>The experiment was run on the SurveyGizmo online questionnaire platform. Before beginning the experiment, subjects provided informed consent and were presented with a short training session showing the types of questions they would encounter. Subjects were instructed that &#8216;different&#8217; responses indicated subtle differences in pronunciation, and not whole word, sound, or speaker differences. Subjects were also instructed to use headphones in a setting free of distractions.</p>
</sec>
<sec>
<title>2.4. Stimuli selection and manipulation</title>
<p>The target word list included 96 unique words beginning with one of six voiced and voiceless stop consonants. All words were a single syllable with no initial consonant clusters, of the form /CV(C)(C)/. Each of these words occurred with a reversed-voicing minimal pair counterpart&#8212;e.g., base/pace, two/do, goat/coat&#8212;which ensured that the phonological environments after the initial consonant was controlled for voiced versus voiceless stimuli as a whole (e.g., if there would happen to be an effect of the following vowel). Each of 96 words was presented to each subject once, in which case it was either the same when repeated, or different, with a raised f0 contour. Two groups of 50 subjects each were presented with exactly half of the stimuli in order to have &#8216;different&#8217; stimuli for each of the 96 target words. Subjects in Group A, for example, heard &#8220;buy&#8221; <italic>different</italic> and &#8220;back&#8221; the <italic>same</italic>, whereas subjects in Group B heard &#8220;buy&#8221; the <italic>same</italic> and &#8220;back&#8221; <italic>different</italic>.</p>
<p>The stimuli were produced by a phonetically-trained male in his 30s, recorded in a quiet location with a Zoom H4n Handy recorder. The voiceless stimuli were produced with aspiration, while the voiced stimuli were produced with voicing during the stop closure. The Praat Manipulate tool was used to alter the pitch contours for the stimuli. A neutral f0 contour was extracted from a vowel-initial word and was applied to all the base utterances. For the &#8216;different&#8217; repetitions, the f0 was manipulated to begin 18 Hz higher than the base utterance and gradually slope downward, reaching the same pitch as the base utterance after about 125 ms. The base f0 contour and the raised different f0 contour are shown in Figure <xref ref-type="fig" rid="F5">5</xref>.</p>
<fig id="F5">
<label>Figure 5</label>
<caption>
<p>Manipulated F0 contours of stimuli.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75834/"/>
</fig>
</sec>
</sec>
<sec>
<title>3. Results</title>
<p>Data was collected from a total of 100 subjects who completed the experiment. All data was kept, even in cases when the subjects did not do better than chance at the discrimination task. Thus, with 96 AX discrimination stimuli, there were a total of 9600 responses. Of these responses, 58.4% were correct (in noting either sameness or difference), with 39.9% incorrect responses, and 1.6% stimuli left unanswered. This suggests some difficulty with the discrimination task, yet clearly indicates that given the number of subjects and trials, subjects did significantly better than chance (50%, with a binomial test yielding <italic>p</italic> &lt; 0.0001) and thus fully understood the nature of the experiment as a whole. There was also a clear response bias towards believing the stimuli sounded the &#8216;same&#8217; when repeated, evidenced by a 70.25% success rate for the &#8216;same&#8217; tokens compared to only 46.6% success for the &#8216;different&#8217; tokens. This again suggests the acoustic difference in the repeated word was subtle.</p>
<p>By examining only the &#8216;different&#8217; stimuli, we can determine whether subjects were more likely to notice the difference for voiced consonant stimuli&#8212;as is hypothesized&#8212;than for voiceless consonant stimuli. Here we see a clear bias, as predicted.</p>
<p>Looking at just the 4800 &#8216;different&#8217; stimuli we can begin to assess whether subjects were more likely to notice an increased pitch contour on voiced stimuli as opposed to voiceless ones. Subjects showed a 52.9% success rate in noticing the raised f0 contour on the voiced stimuli (1269/2400) compared to only a 40.4% success rate in noticing the raised f0 contour on the voiceless stimuli (970/2400). This demonstrates a much higher success rate in noticing the raised f0 contour for the voiced stimuli, where such a contour would not arise from coarticulation.</p>
<sec>
<title>3.1. D&#8217; statistical analysis</title>
<p>The sensitivity index, <italic>d&#8217;</italic>, was calculated to assess the statistic significance of the data. This statistical measure is useful for determining a subject&#8217;s ability to perceive similarity or difference while taking into account the possibility of response biases&#8212;for example, subjects who primarily think everything sounds the &#8216;same.&#8217; Thus, this statistic takes into account not only a subject&#8217;s &#8216;hits&#8217;&#8212;correctly identifying when the word repetition is different&#8212;but also &#8216;false alarms,&#8217; where the subject believes the stimuli sounded different when they were in fact the same. Thus, a subject with a 100% hit rate but also a 100% false alarm rate would have a very low <italic>d&#8217;</italic> score, whereas a 100% hit rate and a 0% false alarm rate would result in a high score, indicating a high level of sensitivity in detecting the acoustic differences in the stimuli (<xref ref-type="bibr" rid="B28">MacMillan &amp; Creelman, 2005</xref>). Two <italic>d&#8217;</italic> values were calculated for each subject, one quantifying sensitivity towards differences in the voiced stimuli and the other for the voiceless stimuli, using the same-different <italic>d&#8217;</italic> equation (rather than &#8216;yes-no&#8217;) via the sensR package in R. Thus, we can compare over all subjects to see if their <italic>d&#8217;</italic> scores are significantly higher for the voiced tokens, as is hypothesized.</p>
<p>Additionally, since <italic>d&#8217;</italic> is less accurate for very high hit and false alarm rates approaching 100% or 0% due to resulting in infinite <italic>z-</italic>scores (<xref ref-type="bibr" rid="B40">Stanislaw &amp; Todorov, 1999</xref>), I followed a similar method to the log-linear approach detailed in Hautus (<xref ref-type="bibr" rid="B9">1995</xref>). In order to avoid 0% and 100% rates, a value of 1 was added to each of the false alarm, correct rejection, miss, and hit totals for each subject. Thus, a perfect hit rate of 24/24 and 0/24 misses would be corrected to 25/26 and 1/26 respectively.</p>
<p><italic>D&#8217;</italic> was calculated for all 100 subjects in both the voiced and voiceless stimuli conditions. The mean <italic>d&#8217;</italic> for the voiced stimuli over all subjects was 1.59, whereas the mean for the voiceless stimuli was only 0.96. Subjects also showed a bias towards responding &#8216;same,&#8217; rather than &#8216;different,&#8217; reflected in a criteria location value, <italic>c</italic>, of 0.5938, the positive value here indicating a higher miss rate than false alarm rate. In order to assess statistical significance, a paired, two-tailed t-test was conducted comparing the individual voiced versus voiceless stimuli scores for each of the 100 subjects. The results show subjects&#8217; <italic>d&#8217;</italic> for the voiced stimuli was significantly higher than the <italic>d&#8217;</italic> for voiceless stimuli (<italic>p</italic> &lt; 0.0001, <italic>t</italic> = 6.37, <italic>df</italic> = 100), corroborating the assessment that subjects more readily could perceive the raised pitch contour on voiced rather than voiceless stimuli. The scores for all subjects are shown below in Figure <xref ref-type="fig" rid="F6">6</xref>, with <italic>d&#8217;</italic> values for voiceless stimuli along the y-axis and values for voiced stimuli along the x-axis. Each dot represents one of the 100 participants, and the diagonal line represents an equal <italic>d&#8217;</italic> score for both the voiced and voiceless stimuli&#8212;thus those falling above the line demonstrated better discrimination of the voiceless tokens, while those falling below the line demonstrated better discrimination of the voiced tokens. The greater number of dots below the line reflects the bias towards better discrimination of the voiced tokens among individual subjects.</p>
<fig id="F6">
<label>Figure 6</label>
<caption>
<p><italic>d&#8217;</italic> scores for voiced and voiceless stimuli over all subjects. Each dot represents a single subject.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75835/"/>
</fig>
</sec>
<sec>
<title>3.2. Principal Components Analysis</title>
<p>A post-hoc Principal Components Analysis (PCA) was also conducted in order to determine if subjects used different listening and/or response strategies. The PCA was run using the prcomp function in R. The model included twelve variables, which were the total number of correct responses (out of eight) for each of the six phoneme stimuli (/p t k b d g/) repeated either the same or different (6 phonemes &#215; 2 conditions). The results revealed just two principle components that accounted for more than 5% of the data.</p>
<p>As shown in the biplot in Figure <xref ref-type="fig" rid="F7">7</xref>, PC1 explains 57.1% of the variation in the data, whereas PC2 accounts for 18.4%. The red arrows show the contribution of each variable to the two PCs. For example, higher accuracy of the &#8216;different&#8217; stimuli led to a higher PC1 score (thus the leftward points of the &#8216;same&#8217; variables and the rightward points of the &#8216;different&#8217; variables). Thus, we can conclude that PC1 is the component indicating either a &#8216;same&#8217; or &#8216;different&#8217; response bias. For example, subject 6 responded &#8216;different&#8217; to all 96 stimuli, whereas subject 10 responded &#8216;same&#8217; to all but one of the 96 stimuli.</p>
<fig id="F7">
<label>Figure 7</label>
<caption>
<p>Principal Components Analysis of subjects.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6278/file/75836/"/>
</fig>
<p>Higher accuracy of <italic>all</italic> stimuli contributed to a higher PC2 score, suggesting that PC2 correlates with how accurately the subjects responded overall. For example, subject 8, with the highest PC2 score, correctly discriminated 89 of the 96 stimuli. The overall triangular shape of the subjects&#8217; spread is indicative of the fact that the ones who did not have a strong &#8216;same&#8217; or &#8216;different&#8217; response bias usually did better overall. Further inspecting PC2, we see that the three voiced &#8216;different&#8217; variables contributed more to a higher PC2 since subjects were more successful in discriminating those stimuli (thus the higher upward tilt of those arrows).</p>
</sec>
</sec>
<sec>
<title>4. Discussion</title>
<p>The results show that subjects were better at discriminating target word stimuli when the only acoustic difference was variation that would <italic>not</italic> be an expected result of coarticulation, as compared to variation that would be expected to result from coarticulation. That is to say, subjects more readily noticed, stored, and/or maintained a raised f0 contour in memory following <italic>voiced</italic> consonants, but did not do this to the same extent following <italic>voiceless</italic> consonants, where such coarticulation would be expected. The interpretation of these results suggests consequences for both our current knowledge of exemplar theory as well as sound change.</p>
<sec>
<title>4.1. Consequences for exemplar theory</title>
<p>The main findings from this study suggest that either some acoustic detail is filtered from exemplar memory or that certain details fade more rapidly from those exemplars. This departs from some other exemplar models in suggesting some degree of abstraction occurs in the process of storing exemplars. For example, Johnson (<xref ref-type="bibr" rid="B17">1997b</xref>) suggests that the entire auditory spectrum is stored in memory. The current findings suggest that, contrary to this, that not all information is stored in its raw, veridical form. However, Johnson&#8217;s assertion cannot be ruled out entirely. It could be the case that the raw auditory spectrum is briefly stored in a sort of working memory before being committed to exemplar storage. All that can be deduced for sure is that different features of the auditory signal, particularly those which are more predictable in some way, fade from exemplar memory more quickly. Johnson&#8217;s (<xref ref-type="bibr" rid="B16">1997a</xref>) suggestion of reduced exemplars that store only formant information also requires some amendment. It is unclear that there is any automatic means of compressing the auditory data for storage in memory; rather, the listener&#8217;s attention guides what is stored, and details that are less predictable in a particular context are more likely to be committed to longer term storage of the veridical details. Generally speaking, these findings align well with hybrid exemplar models (e.g., <xref ref-type="bibr" rid="B36">Pierrehumbert, 2002</xref>), which suggest listeners store both exemplars and abstract representations in memory. However, it remains unclear whether both types of exemplars exist, or if instead exemplars are typically a patchwork quilt of veridical and abstracted information. While further study is needed to understand the relationship between veridical and abstracted detail in memory, it seems probable that details of individual exemplars will continue to fade from memory and become abstracted over time. At some point, the memory trace may primarily contain phonemic information, or further abstract to represent mere words or ideas. However, this does not preclude the possibility that individual exemplar clouds&#8212;containing thousands of traces of varying degrees of abstraction&#8212;may be linked to some sort of fully abstract representation of the word.</p>
<p>The current findings also fit well with Goldinger&#8217;s (<xref ref-type="bibr" rid="B7">2007</xref>) statement that &#8220;each stored &#8216;exemplar&#8217; is actually a product of perceptual input combined with prior knowledge&#8221; (p. 50), as well as Hawkins&#8217; (<xref ref-type="bibr" rid="B10">2003</xref>, <xref ref-type="bibr" rid="B11">2010</xref>) observations that the speech signal is processed only to the extent needed for extracting linguistic meaning. This also concurs with Goldinger &amp; Azuma&#8217;s (<xref ref-type="bibr" rid="B8">2003</xref>) application of Adaptive Resonance Theory (ART), which suggests that there is no fixed unit of speech perception&#8212;speakers adaptively process speech units in whatever way achieves the quickest comprehension of the linguistic meaning of the speech signal. We could further apply the current findings to this model to suggest that the details processed in speech perception are what end up being encoded in the memory traces themselves. Ultimately, acoustic information that is predictable based on coarticulation or context (such as found in <xref ref-type="bibr" rid="B29">Manker, 2019</xref>) may not survive in the exemplar.</p>
<p>It should also be noted that the current findings do not suggest that any automatic process of abstraction must act in transforming the speech signal before word recognition, as is typical in models of speech perception including normalization (<xref ref-type="bibr" rid="B3">Gerstman, 1968</xref>; <xref ref-type="bibr" rid="B45">Tranm&#252;ller, 1981</xref>, etc.). Rather, those details that are particularly predictable and/or redundant in speech perception may be most likely to fade from exemplar memory. In fact, some subphonemic information is often facilitative in speech recognition (<xref ref-type="bibr" rid="B16">Johnson, 1997a</xref>), and in such cases, I would expect these details would more likely survive in exemplar memory. Further study will examine this question more closely.</p>
<p>While the results strongly demonstrate an effect of perceptual salience of coarticulatory details in a given phonetic context, an alternative analysis could challenge whether the observed perceptual bias is relevant to exemplar storage at all. For example, perhaps listeners at some point in the experiment became aware that &#8216;different&#8217; always meant a raised pitch contour, at which point they began to ignore the initial utterance and only focus on the repetition. In this case, the difference in perceptual salience was all that motivated the observed bias, with no bias in what was originally stored or maintained in exemplar memory. An experimental design including initial utterances with raised pitch which is then repeated would be able to rule out or confirm this possible explanation. However, I believe this alternative account is unlikely, primarily due to the inclusion of the filler &#8216;word-identification&#8217; stimuli. In these cases, no word was repeated at all, so it encouraged listeners to pay close attention to the initial utterance of the word and not only its (possible) repetition. Additionally, the dual tasks would likely distract from subjects&#8217; attempts to determine any patterns in the repetitions. In any case, further research can explore and disambiguate this competing interpretation.</p>
</sec>
<sec>
<title>4.2. Relation of findings to compensation for coarticulation</title>
<p>Compensation for coarticulation is a phenomenon whereby listeners perceptually &#8216;undo&#8217; coarticulatory effects of neighboring sounds in order to determine the intended underlying segments. This was famously observed in Mann and Repp (<xref ref-type="bibr" rid="B30">1980</xref>), in which subjects were more likely to perceive an acoustically ambiguous fricative as the sound [s] rather than [&#643;] following [u], arguably due to the coarticulatory effect caused by the rounded vowel [u], which tends to cause the neighboring sounds to lower in frequency. The mechanics of this process suggest something similar to the effect found in the current paper&#8212;that listeners perceptually remove an initial pitch raise following voiceless sounds as an expected effect of coarticulation, thus stripping away this detail in exemplar memory. However, Holt, Lotto, and Kluender (<xref ref-type="bibr" rid="B12">2001</xref>) find sensitivity to the coarticulatory relationship between F0 and voicing in Japanese quail, suggesting awareness of this relationship is not rooted in human speech perception. As a result, it is not clear if the phenomenon observed in the present paper is the result of compensation for coarticulation or a distinct phenomenon rooted in more general acoustic predictability, though the perceptual consequences may be quite similar. Future research and analysis are needed to investigate the relationship of predictability-modulated acoustic awareness and compensation for coarticulation.</p>
</sec>
<sec>
<title>4.3. Consequences for sound change</title>
<p>Ohala&#8217;s (<xref ref-type="bibr" rid="B34">1981</xref>, <xref ref-type="bibr" rid="B35">1983</xref>) account of perceptual correction, and other studies of compensation for coarticulation (<xref ref-type="bibr" rid="B47">Yu, 2010</xref>; <xref ref-type="bibr" rid="B48">Yu, Abrego-Collier, &amp; Sonderegger, 2013</xref>), have considered the role of perceptual biases in sound change. For example, Ohala (<xref ref-type="bibr" rid="B34">1981</xref>) claims that hypocorrection is one source of sound change&#8212;when speakers notice certain coarticulatory details but do not attribute them to their phonological environment. Yu (<xref ref-type="bibr" rid="B47">2010</xref>) found that female subjects with low Autism Quotient scores demonstrated less compensation for coarticulation, being more likely to notice certain articulatory effects (e.g., hearing [&#643;] before [u] instead of [s] and not attributing the lowered fricative frequencies to coarticulation). The results of the current study could be applied to either of these accounts, following the proposal that such &#8216;misperception&#8217; results from encoding predictable coarticulatory effects in exemplar memory.</p>
<p>The current results make some additional predictions, however. New variation that is phonologically predictable, such as coarticulation, even when exaggerated a bit beyond what listeners are used to hearing, is more likely to evade notice and fail to be retained in memory, at least shortly after perceiving these details. However, new variation that is phonologically unconditioned would be more likely to be stored and maintained in exemplar memory. If, following the phonetic accommodation paradigm, we assume that new exemplar traces inform future productions, then we might expect conditioned versus unconditioned sound changes to spread differently within a language. For example, a conditioned change like the nasalization of vowels before nasal consonants (followed by their eventual loss) should more likely evade the notice of listeners since it is a predictable result of coarticulation, though a change of this sort is under the constant articulatory pressure that causes coarticulation in the first place. On the other hand, an unconditioned change, like a chain shift causing /p t k/ to become /p<sup>h</sup> t<sup>h</sup> k<sup>h</sup>/ in all phonological environments should be more likely to be encoded into listeners&#8217; exemplar memories since the change is not phonologically predictable. However, whereas we might predict its spread from speaker to speaker may happen more rapidly, it is not clear how strong the motivation is to begin in the first place, since unconditioned changes may be influenced more by phonological considerations (e.g., pressures within the sound system) rather than articulatory pressure.</p>
<p>Additionally, it is not clear how the effect observed in this study would be maintained over longer periods of time, necessary for eventually permanent changes in a language. For example, predictable detail, while stored in memory less faithfully and possibly abstracted in some way, may actually be retained for longer, whereas the more veridical memories of unpredictable memory may fade more quickly. In any case, the results may suggest some differences in the way that conditioned and unconditioned sound changes spread, though much more work is needed to understand the relevance, if any, of predictability-based perceptual biases in exemplar storage on sound change.</p>
</sec>
<sec>
<title>4.4. Future research</title>
<p>Several important questions remain in order to understand the nature and contents of exemplars. First of all, it is necessary to continue to survey how different acoustic cues are stored in exemplar memory, and the various perceptual biases that facilitate or impede the storage of various auditory information. One question of particular interest will be whether or not certain coarticulatory cues are in fact stored in memory. While I have suggested that predictable coarticulatory cues may be ignored and perceptually filtered, there is a lot of research showing that coarticulation can be used to facilitate speech recognition (<xref ref-type="bibr" rid="B16">Johnson, 1997a</xref>; <xref ref-type="bibr" rid="B1">Beddor, Krakow, &amp; Lindemann, 2001</xref>). Beddor et al., for example, state that &#8220;listeners use coarticulatory variation as information about sounds that are further up or down the speech stream&#8221; (p. 56). If coarticulation aids in speech recognition in this way, following my previous proposal that details that are used in speech recognition will more likely survive in exemplar memory, this should result in better storage of these details. However, it is unclear how the events of auditory storage and maintenance unfold. For example, perhaps once phoneme or word recognition occurs, such coarticulatory details are rapidly lost from memory or abstracted. Alternatively, there could be some difference in anticipatory versus confirmatory coarticulation. In the present study, the f0 modulation occurred after the voicing and VOT cues had provided ample evidence as to the initial sounds in the target words (e.g., &#8216;bath,&#8217; &#8216;path,&#8217; etc.). The f0 cue only served as an expected confirmation. On the other hand, nasality on a vowel that cues an upcoming nasal consonant occurs at a point in time when the listener does not know the upcoming sound (such as is shown for English by <xref ref-type="bibr" rid="B24">Lahiri &amp; Marslen-Wilson, 1991</xref>). Thus, anticipatory coarticulation of this sort might be more faithfully stored and maintained in memory.</p>
<p>Further research should also consider additional aspects of predictability and expectation. For example, predictability of a person&#8217;s voice may lead to lower awareness of certain acoustic information&#8212;more acoustic information might be stored when listening to different speakers produce stimuli, similar to the stronger imitative effect found in Goldinger (<xref ref-type="bibr" rid="B6">1998</xref>) when subjects heard multiple model speakers. In addition to the contextual predictability effect found in Manker (<xref ref-type="bibr" rid="B29">2019</xref>), which was based on word priming, we could consider whether general situational context also results in lower attention and less faithful storage of the auditory signal. For example, if objects are visually presented before they are referred to, will listeners store less auditory detail of these words? Finally, we might consider how the predictability of semantic context interacts with other forms of predictability, such as phonologically predictable coarticulation.</p>
<p>Lastly, further work in the phonetic accommodation paradigm will complement the findings of the current study. For example, as in Manker (<xref ref-type="bibr" rid="B29">2019</xref>), should we expect to find greater accommodation of predictable coarticulation compared to unpredictable phonetic variation? Zellou, Scarborough, and Nielsen (<xref ref-type="bibr" rid="B50">2016</xref>), for example, did in fact find imitation of &#8216;hyper-nasalized&#8217; vowel coarticulation, such that speakers did in fact increase their own vowel nasality after hearing greater vowel nasality produced. From this study alone, it is not clear whether such imitation of coarticulation is weaker in magnitude than imitation of unconditioned acoustic variation. Secondly, this is again a case of an anticipatory coarticulation effect, so would the effect be weaker for perseverative coarticulation? Further research will help to interpret the growing body of literature in phonetic accommodation as a whole and its relevance in the storage and maintenance of detail in exemplar memory.</p>
</sec>
</sec>
<sec>
<title>5. Conclusion</title>
<p>The present study addresses the question of whether predictable coarticulatory detail is stored and maintained in exemplar memory to the same degree as unpredictable acoustic variation. The results of an AX discrimination task show that subjects did a significantly better job at discriminating tokens that differed in phonologically unpredictable ways as opposed to those that merely displayed expected coarticulatory detail. This suggests some degree of filtering or abstraction occurs in exemplar storage and is modulated by the predictability of the variation. Future research will continue to address the phenomenon of predictability and expectation, how it shapes the contents of exemplars, and its relevance in sound change.</p>
</sec>
<sec sec-type="supplementary-material">
<title>Additional File</title>
<p>The additional file for this article can be found as follows:</p>
<supplementary-material id="S1" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.240.s1">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-11-240-s1.pdf">labphon-11-240-s1.pdf</inline-supplementary-material>]-->
<label>Appendix</label>
<caption>
<p>List of stimuli used for this experiment. This includes the target words, which included voiced-voiceless minimal pairs, as well as fillers, which were single syllable words with no other phonological restrictions. DOI: <uri>https://doi.org/10.5334/labphon.240.s1</uri></p>
</caption>
</supplementary-material>
</sec>
</body>
<back>
<sec>
<title>Competing Interests</title>
<p>The author has no competing financial, professional, or personal interests that might have affected the objectivity or integrity of this publication.</p>
</sec>
<ref-list>
<ref id="B1"><label>1</label><mixed-citation publication-type="book"><string-name><surname>Beddor</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Krakow</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Lindemann</surname>, <given-names>S.</given-names></string-name> (<year>2001</year>). <chapter-title>Patterns of perceptual compensation and their phonological consequences</chapter-title>. In <string-name><given-names>E.</given-names> <surname>Hume</surname></string-name> &amp; <string-name><given-names>K.</given-names> <surname>Johnson</surname></string-name> (Eds.), <source>The Role of Speech Perception in Phonology</source> (pp. <fpage>55</fpage>&#8211;<lpage>78</lpage>). <publisher-loc>San Diego</publisher-loc>: <publisher-name>Academic Press</publisher-name>.</mixed-citation></ref>
<ref id="B2"><label>2</label><mixed-citation publication-type="book"><string-name><surname>Chomsky</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Halle</surname>, <given-names>M.</given-names></string-name> (<year>1968</year>). <source>The Sound Pattern of English</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Harper &amp; Row</publisher-name>.</mixed-citation></ref>
<ref id="B3"><label>3</label><mixed-citation publication-type="journal"><string-name><surname>Gerstman</surname>, <given-names>L.</given-names></string-name> (<year>1968</year>). <article-title>Classification of self-normalized vowels</article-title>. <source>IEEE Transactions on Audio and Electroacoustics, AU-16</source> (pp. <fpage>78</fpage>&#8211;<lpage>80</lpage>). DOI: <pub-id pub-id-type="doi">10.1109/TAU.1968.1161953</pub-id></mixed-citation></ref>
<ref id="B4"><label>4</label><mixed-citation publication-type="journal"><string-name><surname>Gill</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>Gleason</surname>, <given-names>H.</given-names></string-name> (<year>1972</year>). <article-title>The salient features of the Punjabi language</article-title>. <source>Pakha Sanjam</source>, <volume>4</volume>, <fpage>1</fpage>&#8211;<lpage>3</lpage>.</mixed-citation></ref>
<ref id="B5"><label>5</label><mixed-citation publication-type="webpage"><string-name><surname>Goldinger</surname>, <given-names>S. D.</given-names></string-name> (<year>1996</year>). <article-title>Words and voices: Episodic traces in spoken word identification and recognition memory</article-title>. <source>Journal of Experimental Psychology: Learning, Memory, and Cognition</source>, <volume>22</volume>, <fpage>1166</fpage>&#8211;<lpage>1183</lpage>. Retrieved from: <uri>http://citeseerx.ist.psu.edu/viewdoc/download?doi=10.1.1.381.4638&amp;rep=rep1&amp;type=pdf</uri>. DOI: <pub-id pub-id-type="doi">10.1037/0278-7393.22.5.1166</pub-id></mixed-citation></ref>
<ref id="B6"><label>6</label><mixed-citation publication-type="journal"><string-name><surname>Goldinger</surname>, <given-names>S. D.</given-names></string-name> (<year>1998</year>). <article-title>Echoes of echoes? An episodic theory of lexical access</article-title>. <source>Psychological Review</source>, <volume>105</volume>(<issue>2</issue>), <fpage>251</fpage>&#8211;<lpage>279</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0033-295X.105.2.251</pub-id></mixed-citation></ref>
<ref id="B7"><label>7</label><mixed-citation publication-type="confproc"><string-name><surname>Goldinger</surname>, <given-names>S. D.</given-names></string-name> (<year>2007</year>). <article-title>A complementary-systems approach to abstract and episodic speech perception</article-title>. <conf-name>Proceedings of the 17th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>49</fpage>&#8211;<lpage>54</lpage>). Retrieved from: <uri>http://icphs2007.de/conference/Papers/1781/1781.pdf</uri></mixed-citation></ref>
<ref id="B8"><label>8</label><mixed-citation publication-type="journal"><string-name><surname>Goldinger</surname>, <given-names>S. D.</given-names></string-name>, &amp; <string-name><surname>Azuma</surname>, <given-names>T.</given-names></string-name> (<year>2003</year>). <article-title>Puzzle-solving science: The quixotic quest for units in speech perception</article-title>. <source>Journal of Phonetics</source>, <volume>31</volume>, <fpage>305</fpage>&#8211;<lpage>320</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/S0095-4470(03)00030-5</pub-id></mixed-citation></ref>
<ref id="B9"><label>9</label><mixed-citation publication-type="journal"><string-name><surname>Hautus</surname>, <given-names>M. J.</given-names></string-name> (<year>1995</year>). <article-title>Corrections for extreme proportions and their biasing effects on estimated values of <italic>d&#8217;</italic></article-title>. <source>Behavior Research Methods, Instruments, &amp; Computers</source>, <volume>27</volume>, <fpage>46</fpage>&#8211;<lpage>51</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03203619</pub-id></mixed-citation></ref>
<ref id="B10"><label>10</label><mixed-citation publication-type="journal"><string-name><surname>Hawkins</surname>, <given-names>S.</given-names></string-name> (<year>2003</year>). <article-title>Roles and representations of systematic fine phonetic detail in speech understanding</article-title>. <source>Journal of Phonetics</source>, <volume>31</volume>(<issue>3&#8211;4</issue>), <fpage>373</fpage>&#8211;<lpage>405</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2003.09.006</pub-id></mixed-citation></ref>
<ref id="B11"><label>11</label><mixed-citation publication-type="journal"><string-name><surname>Hawkins</surname>, <given-names>S.</given-names></string-name> (<year>2010</year>). <article-title>Phonetic variation as communicative system: Perception of the particular and the abstract</article-title>. <source>Laboratory Phonology</source>, <volume>10</volume>, <fpage>479</fpage>&#8211;<lpage>510</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/9783110224917.5.479</pub-id></mixed-citation></ref>
<ref id="B12"><label>12</label><mixed-citation publication-type="journal"><string-name><surname>Holt</surname>, <given-names>L. L.</given-names></string-name>, <string-name><surname>Lotto</surname>, <given-names>A. J.</given-names></string-name>, &amp; <string-name><surname>Kluender</surname>, <given-names>K. R.</given-names></string-name> (<year>2001</year>). <article-title>Influence of fundamental frequency on stop-consonant voicing perception: A case of learned covariation or auditory enhancement?</article-title> <source>Journal of the Acoustical Society of America</source>, <volume>109</volume>, <fpage>764</fpage>&#8211;<lpage>774</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1339825</pub-id></mixed-citation></ref>
<ref id="B13"><label>13</label><mixed-citation publication-type="journal"><string-name><surname>Hombert</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Ladefoged</surname>, <given-names>P.</given-names></string-name> (<year>1976</year>). <article-title>The effect of aspiration on the fundamental frequency of the following vowel</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>59</volume>, <fpage>S72</fpage> (abstract). DOI: <pub-id pub-id-type="doi">10.1121/1.2002863</pub-id></mixed-citation></ref>
<ref id="B14"><label>14</label><mixed-citation publication-type="webpage"><string-name><surname>Hombert</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Ohala</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Ewan</surname>, <given-names>W.</given-names></string-name> (<year>1979</year>). <article-title>Phonetic explanations for the development of tones</article-title>. <source>Language</source>, <volume>55</volume>, <fpage>37</fpage>&#8211;<lpage>58</lpage>. Retrieved from: <uri>http://linguistics.berkeley.edu/~ohala/papers/phonet_expl_tones.pdf</uri>. DOI: <pub-id pub-id-type="doi">10.2307/412518</pub-id></mixed-citation></ref>
<ref id="B15"><label>15</label><mixed-citation publication-type="journal"><string-name><surname>House</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Fairbanks</surname>, <given-names>G.</given-names></string-name> (<year>1953</year>). <article-title>The influence of consonant environment upon the secondary acoustical characteristics of vowels</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>10</volume>, <fpage>105</fpage>&#8211;<lpage>113</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1906982</pub-id></mixed-citation></ref>
<ref id="B16"><label>16</label><mixed-citation publication-type="webpage"><string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> (<year>1997a</year>). <chapter-title>Speech perception without speaker normalization: An exemplar model</chapter-title>. In <string-name><surname>Johnson</surname></string-name> &amp; <string-name><surname>Mullennix</surname></string-name> (Eds.), <source>Talker Variability in Speech Processing</source> (pp. <fpage>145</fpage>&#8211;<lpage>165</lpage>). <publisher-loc>San Diego</publisher-loc>: <publisher-name>Academic Press</publisher-name>. Retrieved from: <uri>http://linguistics.berkeley.edu/~kjohnson/papers/SpeechPerceptionWithoutSpeakerNormalization.pdf</uri></mixed-citation></ref>
<ref id="B17"><label>17</label><mixed-citation publication-type="webpage"><string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> (<year>1997b</year>). <chapter-title>The auditory/perceptual basis for speech segmentation</chapter-title>. <source>OSU Working Papers in Linguistics</source>, <volume>50</volume>, <fpage>101</fpage>&#8211;<lpage>113</lpage>. <publisher-loc>Columbus, Ohio</publisher-loc>. Retrieved from: <uri>http://linguistics.berkeley.edu/~kjohnson/papers/Johnson1997.pdf</uri></mixed-citation></ref>
<ref id="B18"><label>18</label><mixed-citation publication-type="book"><string-name><surname>Johnson</surname>, <given-names>K.</given-names></string-name> (<year>2007</year>). <chapter-title>Decisions and mechanisms in exemplar-based phonology</chapter-title>. In <string-name><given-names>M. J.</given-names> <surname>Sol&#233;</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Beddor</surname></string-name> &amp; <string-name><given-names>M.</given-names> <surname>Ohala</surname></string-name> (Eds.), <source>Experimental Approaches to Phonology. In Honor of John Ohala</source> (pp. <fpage>25</fpage>&#8211;<lpage>40</lpage>). <publisher-name>Oxford University Press</publisher-name>.</mixed-citation></ref>
<ref id="B19"><label>19</label><mixed-citation publication-type="journal"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name> (<year>1989</year>). <article-title>The effect of macroscopic context on consonantal perturbations of fundamental frequency</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>85</volume>, <fpage>S149</fpage> (abstract). DOI: <pub-id pub-id-type="doi">10.1121/1.2026802</pub-id></mixed-citation></ref>
<ref id="B20"><label>20</label><mixed-citation publication-type="book"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name> (<year>2011</year>). <chapter-title>Tonogenesis</chapter-title>. In <string-name><given-names>M.</given-names> <surname>van Oostendorp</surname></string-name>, <string-name><given-names>J. Ewen</given-names> <surname>Colin</surname></string-name>, <string-name><given-names>E.</given-names> <surname>Hume</surname></string-name> &amp; <string-name><given-names>K.</given-names> <surname>Rice</surname></string-name> (Eds.), <source>The Blackwell companion to phonology</source> (pp. Chapter 97). <publisher-loc>Malden, MA &amp; Oxford</publisher-loc>. <publisher-name>Wiley-Blackwell</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1002/9781444335262.wbctp0097</pub-id></mixed-citation></ref>
<ref id="B21"><label>21</label><mixed-citation publication-type="journal"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Diehl</surname>, <given-names>R.</given-names></string-name> (<year>1994</year>). <article-title>Phonetic knowledge</article-title>. <source>Language</source>, <volume>70</volume>, <fpage>419</fpage>&#8211;<lpage>454</lpage>. DOI: <pub-id pub-id-type="doi">10.1353/lan.1994.0023</pub-id></mixed-citation></ref>
<ref id="B22"><label>22</label><mixed-citation publication-type="journal"><string-name><surname>Kingston</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Diehl</surname>, <given-names>R. L.</given-names></string-name>, <string-name><surname>Kirk</surname>, <given-names>C. J.</given-names></string-name>, &amp; <string-name><surname>Castleman</surname>, <given-names>W. A.</given-names></string-name> (<year>2008</year>). <article-title>On the internal perceptual structure of distinctive features: The [voice] contrast</article-title>. <source>Journal of Phonetics</source>, <volume>36</volume>, <fpage>28</fpage>&#8211;<lpage>54</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2007.02.001</pub-id></mixed-citation></ref>
<ref id="B23"><label>23</label><mixed-citation publication-type="book"><string-name><surname>Klatt</surname>, <given-names>D. H.</given-names></string-name> (<year>1979</year>). <chapter-title>Speech perception: A model of acoustic-phonetic analysis and lexical access</chapter-title>. In <string-name><given-names>R. A.</given-names> <surname>Cole</surname></string-name> (ed.), <source>Perception and production of fluent speech</source> (pp. <fpage>243</fpage>&#8211;<lpage>288</lpage>). <publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1016/S0095-4470(19)31059-9</pub-id></mixed-citation></ref>
<ref id="B24"><label>24</label><mixed-citation publication-type="journal"><string-name><surname>Lahiri</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Marslen-Wilson</surname>, <given-names>W.</given-names></string-name> (<year>1991</year>). <article-title>The mental representation of lexical form: A phonological approach to the recognition lexicon</article-title>. <source>Cognition</source>, <volume>38</volume>(<issue>3</issue>), <fpage>245</fpage>&#8211;<lpage>294</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/0010-0277(91)90008-R</pub-id></mixed-citation></ref>
<ref id="B25"><label>25</label><mixed-citation publication-type="journal"><string-name><surname>Lehiste</surname>, <given-names>I.</given-names></string-name>, &amp; <string-name><surname>Peterson</surname>, <given-names>G.</given-names></string-name> (<year>1961</year>). <article-title>Some basic considerations in the analysis of intonation</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>33</volume>, <fpage>419</fpage>&#8211;<lpage>425</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.1908681</pub-id></mixed-citation></ref>
<ref id="B26"><label>26</label><mixed-citation publication-type="journal"><string-name><surname>Liberman</surname>, <given-names>A. M.</given-names></string-name>, <string-name><surname>Cooper</surname>, <given-names>F. S.</given-names></string-name>, <string-name><surname>Shankweiler</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Studdert-Kennedy</surname>, <given-names>M.</given-names></string-name> (<year>1967</year>). <article-title>Perception of the speech code</article-title>. <source>Psychological Review</source>, <volume>74</volume>, <fpage>431</fpage>&#8211;<lpage>461</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/h0020279</pub-id></mixed-citation></ref>
<ref id="B27"><label>27</label><mixed-citation publication-type="webpage"><string-name><surname>Luce</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>McLennan</surname>, <given-names>C.</given-names></string-name> (<year>2005</year>). <chapter-title>Spoken word recognition: The challenge of variation</chapter-title>. In <string-name><given-names>D. B.</given-names> <surname>Pisoni</surname></string-name> &amp; <string-name><given-names>R. E.</given-names> <surname>Remez</surname></string-name> (Eds.), <source>Handbook of Speech Perception</source> (pp. <fpage>591</fpage>&#8211;<lpage>609</lpage>). <publisher-loc>Malden, MA</publisher-loc>: <publisher-name>Blackwell</publisher-name>. Retrieved from: <uri>https://pdfs.semanticscholar.org/1eb3/da8316bf9a3804f32c6d1414a68331c7404e.pdf</uri>. DOI: <pub-id pub-id-type="doi">10.1002/9780470757024.ch24</pub-id></mixed-citation></ref>
<ref id="B28"><label>28</label><mixed-citation publication-type="book"><string-name><surname>MacMillan</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Creelman</surname>, <given-names>C.</given-names></string-name> (<year>2005</year>). <source>Detection Theory: A User&#8217;s Guide</source> (<edition>2nd ed.</edition>). <publisher-loc>Mahwah, NJ</publisher-loc>: <publisher-name>Lawrence Erlbaum Associates</publisher-name>.</mixed-citation></ref>
<ref id="B29"><label>29</label><mixed-citation publication-type="journal"><string-name><surname>Manker</surname>, <given-names>J.</given-names></string-name> (<year>2019</year>). <article-title>Contextual Predictability and Phonetic Attention</article-title>. <source>Journal of Phonetics</source>, <volume>75</volume>, <fpage>94</fpage>&#8211;<lpage>112</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2019.05.005</pub-id></mixed-citation></ref>
<ref id="B30"><label>30</label><mixed-citation publication-type="journal"><string-name><surname>Mann</surname>, <given-names>V.</given-names></string-name>, &amp; <string-name><surname>Repp</surname>, <given-names>B.</given-names></string-name> (<year>1980</year>). <article-title>Influence of vocalic context on the perception of the [&#643;]-[s] distinction</article-title>. <source>Perception and Psychophysics</source>, <volume>28</volume>, <fpage>213</fpage>&#8211;<lpage>228</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03204377</pub-id></mixed-citation></ref>
<ref id="B31"><label>31</label><mixed-citation publication-type="journal"><string-name><surname>McLennan</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Luce</surname>, <given-names>P.</given-names></string-name> (<year>2005</year>). <article-title>Examining the Time Course of Indexical Specificity Effects in Spoken Word Recognition</article-title>. <source>Journal of Experimental Psychology: Learning, Memory and Cognition</source>, <volume>31</volume>(<issue>2</issue>), <fpage>306</fpage>&#8211;<lpage>321</lpage>. DOI: <pub-id pub-id-type="doi">10.1037/0278-7393.31.2.306</pub-id></mixed-citation></ref>
<ref id="B32"><label>32</label><mixed-citation publication-type="journal"><string-name><surname>Nielsen</surname>, <given-names>K.</given-names></string-name> (<year>2011</year>). <article-title>Specificity and abstractness of VOT imitation</article-title>. <source>Journal of Phonetics</source>, <volume>39</volume>(<issue>2</issue>), <fpage>132</fpage>&#8211;<lpage>142</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2010.12.007</pub-id></mixed-citation></ref>
<ref id="B33"><label>33</label><mixed-citation publication-type="webpage"><string-name><surname>Nye</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Fowler</surname>, <given-names>C.</given-names></string-name> (<year>2003</year>). <article-title>Shadowing latency and imitation: The effect of familiarity with the phonetic patterning of English</article-title>. <source>Journal of Phonetics</source>, <volume>31</volume>(<issue>1</issue>), <fpage>63</fpage>&#8211;<lpage>79</lpage>. Retrieved from: <uri>http://www.haskins.yale.edu/Reprints/HL1279.pdf</uri>. DOI: <pub-id pub-id-type="doi">10.1016/S0095-4470(02)00072-4</pub-id></mixed-citation></ref>
<ref id="B34"><label>34</label><mixed-citation publication-type="confproc"><string-name><surname>Ohala</surname>, <given-names>J.</given-names></string-name> (<year>1981</year>). <article-title>The listener as the source of sound change</article-title>. In <string-name><given-names>C.</given-names> <surname>Masek</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Hendrick</surname></string-name> &amp; <string-name><given-names>M.</given-names> <surname>Miller</surname></string-name> (Eds.), <conf-name>Papers from the parasession on language and behavior</conf-name> (pp. <fpage>178</fpage>&#8211;<lpage>203</lpage>). <conf-loc>Chicago</conf-loc>: <conf-sponsor>Chicago Linguistics Society</conf-sponsor>. Retrieved from: <uri>http://linguistics.berkeley.edu/~ohala/papers/listener_as_source.pdf</uri></mixed-citation></ref>
<ref id="B35"><label>35</label><mixed-citation publication-type="webpage"><string-name><surname>Ohala</surname>, <given-names>J.</given-names></string-name> (<year>1983</year>). <chapter-title>The origin of sound patterns in vocal tract constraints</chapter-title>. In <string-name><given-names>P.</given-names> <surname>MacNeilage</surname></string-name> (Ed.), <source>The Production of Speech</source> (pp. <fpage>189</fpage>&#8211;<lpage>216</lpage>). <publisher-loc>New York</publisher-loc>: <publisher-name>Springer-Verlag</publisher-name>. Retrieved from: <uri>http://linguistics.berkeley.edu/~ohala/papers/macn83.pdf</uri>. DOI: <pub-id pub-id-type="doi">10.1007/978-1-4613-8202-7_9</pub-id></mixed-citation></ref>
<ref id="B36"><label>36</label><mixed-citation publication-type="book"><string-name><surname>Pierrehumbert</surname>, <given-names>J.</given-names></string-name> (<year>2002</year>). <chapter-title>Word-specific phonetics</chapter-title>. In <string-name><given-names>C.</given-names> <surname>Gussenhoven</surname></string-name> &amp; <string-name><given-names>N.</given-names> <surname>Warner</surname></string-name> (Eds.), <source>Laboratory Phonology VII</source> (pp. <fpage>101</fpage>&#8211;<lpage>139</lpage>). <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1515/9783110197105.101</pub-id></mixed-citation></ref>
<ref id="B37"><label>37</label><mixed-citation publication-type="journal"><string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name> (<year>2016</year>). <article-title>Phonological representation: Beyond abstract versus episodic</article-title>. <source>Annual Review of Linguistics</source>, <volume>2</volume>, <fpage>33</fpage>&#8211;<lpage>52</lpage>. DOI: <pub-id pub-id-type="doi">10.1146/annurev-linguistics-030514-125050</pub-id></mixed-citation></ref>
<ref id="B38"><label>38</label><mixed-citation publication-type="book"><string-name><surname>Ratliff</surname>, <given-names>M.</given-names></string-name> (<year>2015</year>). <chapter-title>Tonoexodus, tonogenesis, and tone change</chapter-title>. In <string-name><given-names>P.</given-names> <surname>Honeybone</surname></string-name> &amp; <string-name><given-names>J.</given-names> <surname>Salmons</surname></string-name> (Eds.), <source>Handbook of Historical Phonology</source> (pp. <fpage>245</fpage>&#8211;<lpage>261</lpage>). <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1093/oxfordhb/9780199232819.013.021</pub-id></mixed-citation></ref>
<ref id="B39"><label>39</label><mixed-citation publication-type="journal"><string-name><surname>Shockley</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Sabadini</surname>, <given-names>L.</given-names></string-name>, &amp; <string-name><surname>Fowler</surname>, <given-names>C.</given-names></string-name> (<year>2004</year>). <article-title>Imitation in shadowing words</article-title>. <source>Perception and Psychophysics</source>, <volume>66</volume>, <fpage>422</fpage>&#8211;<lpage>429</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03194890</pub-id></mixed-citation></ref>
<ref id="B40"><label>40</label><mixed-citation publication-type="journal"><string-name><surname>Stanislaw</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>Todorov</surname>, <given-names>N.</given-names></string-name> (<year>1999</year>). <article-title>Calculation of signal detection theory measures</article-title>. <source>Behavior Research Methods, Instruments, &amp; Computers</source>, <volume>31</volume>, <fpage>137</fpage>&#8211;<lpage>149</lpage>. DOI: <pub-id pub-id-type="doi">10.3758/BF03207704</pub-id></mixed-citation></ref>
<ref id="B41"><label>41</label><mixed-citation publication-type="book"><string-name><surname>Stevens</surname>, <given-names>K.</given-names></string-name> (<year>1972</year>). <chapter-title>The quantal nature of speech: Evidence from articulatory acoustic data</chapter-title>. In: <string-name><surname>David</surname>, <given-names>Denes</given-names></string-name>, (Ed.), <source>Human communication: A unified view</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>McGraw-Hill</publisher-name>.</mixed-citation></ref>
<ref id="B42"><label>42</label><mixed-citation publication-type="book"><string-name><surname>Svantesson</surname>, <given-names>J.-O.</given-names></string-name> (<year>1991</year>). <chapter-title>Hu: A language with unorthodox tonogenesis</chapter-title>. In <string-name><given-names>J.</given-names> <surname>Davidson</surname></string-name> (Ed.), <source>Austroasiatic Languages: Essays in Honour of H. L. Shorto</source> (pp. <fpage>67</fpage>&#8211;<lpage>79</lpage>). <publisher-loc>London</publisher-loc>: <publisher-name>SOAS</publisher-name>.</mixed-citation></ref>
<ref id="B43"><label>43</label><mixed-citation publication-type="book"><string-name><surname>Thurgood</surname>, <given-names>G.</given-names></string-name> (<year>1999</year>). <source>From Ancient Cham to Modern Dialects: Two Hundred Years of Language Contact and Change</source>. <publisher-loc>Honolulu</publisher-loc>: <publisher-name>University of Hawai&#8216;i Press</publisher-name>.</mixed-citation></ref>
<ref id="B44"><label>44</label><mixed-citation publication-type="journal"><string-name><surname>Tilsen</surname>, <given-names>S.</given-names></string-name> (<year>2009</year>). <article-title>Subphonemic and cross-phonemic priming in vowel shadowing: Evidence for the involvement of exemplars in production</article-title>. <source>Journal of Phonetics</source>, <volume>37</volume>(<issue>3</issue>), <fpage>276</fpage>&#8211;<lpage>296</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2009.03.004</pub-id></mixed-citation></ref>
<ref id="B45"><label>45</label><mixed-citation publication-type="journal"><string-name><surname>Tranm&#252;ller</surname>, <given-names>H.</given-names></string-name> (<year>1981</year>). <article-title>Perceptual dimension of openness in vowels</article-title>. <source>Journal of the Acoustic Society of America</source>, <volume>69</volume>, <fpage>1465</fpage>&#8211;<lpage>1475</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.385780</pub-id></mixed-citation></ref>
<ref id="B46"><label>46</label><mixed-citation publication-type="book"><string-name><surname>Wolff</surname>, <given-names>E.</given-names></string-name> (<year>1987</year>). <chapter-title>Consonant-tone interference in Chadic and its implications for a theory of tonogenesis in Afroasiatic</chapter-title>. In <string-name><given-names>D.</given-names> <surname>Barreteau</surname></string-name> (Ed.), <source>Langues et cultures dans le basin du Lac Tchad</source> (pp. <fpage>193</fpage>&#8211;<lpage>216</lpage>). <publisher-loc>Paris</publisher-loc>: <publisher-name>ORSTOM</publisher-name>.</mixed-citation></ref>
<ref id="B47"><label>47</label><mixed-citation publication-type="journal"><string-name><surname>Yu</surname>, <given-names>A.</given-names></string-name> (<year>2010</year>). <article-title>&#8216;Perceptual compensation is correlated with individuals&#8217; &#8220;autistic&#8221; traits: Implications for models of sound change.&#8217; 2010</article-title>. <source>PLoS ONE</source>, <volume>5</volume>(<issue>8</issue>). DOI: <pub-id pub-id-type="doi">10.1371/journal.pone.0011950</pub-id></mixed-citation></ref>
<ref id="B48"><label>48</label><mixed-citation publication-type="journal"><string-name><surname>Yu</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Abrego-Collier</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Sonderegger</surname>, <given-names>M.</given-names></string-name> (<year>2013</year>). <article-title>Phonetic imitation from an individual-difference perspective: Subjective attitude, personality, and &#8216;autistic&#8217; traits</article-title>. <source>PLOS ONE</source>, <volume>8</volume>(<issue>9</issue>), <elocation-id>e74746</elocation-id>. DOI: <pub-id pub-id-type="doi">10.1371/journal.pone.0074746</pub-id></mixed-citation></ref>
<ref id="B49"><label>49</label><mixed-citation publication-type="journal"><string-name><surname>Yu</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Lee</surname>, <given-names>H.</given-names></string-name> (<year>2014</year>). <article-title>The stability of perceptual compensation for coarticulation within and across individuals. A cross-validation study</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>136</volume>(<issue>1</issue>), <fpage>382</fpage>&#8211;<lpage>388</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4883380</pub-id></mixed-citation></ref>
<ref id="B50"><label>50</label><mixed-citation publication-type="journal"><string-name><surname>Zellou</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Scarborough</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Nielsen</surname>, <given-names>K.</given-names></string-name> (<year>2016</year>). <article-title>Phonetic imitation of coarticulatory vowel nasalization</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>140</volume>, <fpage>3560</fpage>&#8211;<lpage>3575</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4966232</pub-id></mixed-citation></ref>
</ref-list>
</back>
</article>