<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.1/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.1" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology: Journal of the Association for Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/labphon.229</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Intonational variation and incrementality in listener judgments of ethnicity</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Holliday</surname>
<given-names>Nicole</given-names>
</name>
<email>nicole.holliday@pomona.edu</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Villarreal</surname>
<given-names>Dan</given-names>
</name>
<xref ref-type="aff" rid="aff-2">2</xref>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>Department of Linguistics and Cognitive Science, Pomona College, Claremont, CA, US</aff>
<aff id="aff-2"><label>2</label>New Zealand Institute of Language, Brain and Behaviour, University of Canterbury, Christchurch, NZ</aff>
<aff id="aff-3"><label>3</label>Department of Linguistics, University of Pittsburgh, US</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2020-04-01">
<day>01</day>
<month>04</month>
<year>2020</year>
</pub-date>
<pub-date pub-type="collection">
<year>2020</year>
</pub-date>
<volume>11</volume>
<issue>1</issue>
<elocation-id>3</elocation-id>
<history>
<date date-type="received" iso-8601-date="2019-09-22">
<day>22</day>
<month>09</month>
<year>2019</year>
</date>
<date date-type="accepted" iso-8601-date="2020-02-20">
<day>20</day>
<month>02</month>
<year>2020</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2020 The Author(s)</copyright-statement>
<copyright-year>2020</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.5334/labphon.229/"/>
<abstract>
<p>The current study examines how listeners make gradient and variable ethnolinguistic judgments in an experimental context where the speaker&#8217;s identity is well-known. It features an open-guise experiment (<xref ref-type="bibr" rid="B43">Soukup, 2013</xref>) that assessed whether sociolinguistic judgments are subject to <italic>incrementality</italic>, with judgments increasing in magnitude as variable stimuli demonstrate more extreme differences. In particular, this task tested whether judgments of President Barack Obama as sounding &#8216;more&#8217; or &#8216;less&#8217; black (e.g., <xref ref-type="bibr" rid="B1">Alim &amp; Smitherman, 2012</xref>) are sensitive to differences in intonation. Half of critical stimuli featured an L+H* pitch accent, which occurs more frequently in African American Language than in Mainstream U.S. English (<xref ref-type="bibr" rid="B15">Holliday, 2016</xref>). Four stimuli apiece were created from these phrases by making each pitch accent more extreme by semitone-based F0 steps. Seventy-nine listeners rated these stimuli via the question, &#8220;How black does Obama sound here?&#8221; Mixed-effects modeling indicated that listeners rated more phonetically extreme L+H* stimuli as sounding blacker, regardless of listener identity. A post-hoc analysis found that listeners attended to different voice quality features in L+H* stimuli. We discuss implications for research in intonation, ethnic identification, incrementality, language attitudes, and sociolinguistic awareness.</p>
</abstract>
<kwd-group>
<kwd>Intonation</kwd>
<kwd>perception</kwd>
<kwd>sociophonetics</kwd>
<kwd>African American Language</kwd>
<kwd>language attitudes</kwd>
<kwd>variation</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>1. Introduction</title>
<p>Recent research in perceptual sociolinguistics has investigated a host of phonetic and phonological variables&#8212;primarily segmental&#8212;to assess the extent to which social meanings are constructed in perception, similar to the way they are constructed in ongoing production. Despite production research in sociolinguistics demonstrating how speakers use intonational variation to index various ethnic identities and social stances (<xref ref-type="bibr" rid="B4">Burdin, 2015</xref>; <xref ref-type="bibr" rid="B15">Holliday, 2016</xref>; <xref ref-type="bibr" rid="B41">Reed, 2016</xref>), there has been a general lack of perceptual research on the social meanings of intonational variables. In addition, while decades of research have demonstrated U.S. listeners&#8217; ability to distinguish African American and white voices (cf. <xref ref-type="bibr" rid="B48">Thomas &amp; Reaser, 2004</xref>), these studies have also revealed challenges inherent in isolating speaker-specific variables that drive ethnic identification (<xref ref-type="bibr" rid="B16">Holliday &amp; Jaggers, 2015</xref>; <xref ref-type="bibr" rid="B39">Purnell, Idsardi, &amp; Baugh, 1999</xref>); indeed, there has been little research on prosody more generally in ethnolinguistic and regional varieties of English (<xref ref-type="bibr" rid="B5">Burdin, Holliday, &amp; Reed, 2018</xref>). In the present study, we address these gaps in research by investigating the extent to which listeners perceive specific aspects of intonational variation as indexes of ethnic identity.</p>
<p>In addition, research in perceptual sociolinguistics has rarely confronted the issue of whether social meanings are <italic>incremental</italic>&#8212;that is, how the social meanings of gradient features are affected by these features&#8217; phonetic shape. Put differently, does a more phonetically extreme token of a socially marked variable correspond to a stronger social meaning? This gap is partially due to the common practice of treating continuous socially marked variables as categorical, such as /&#633;/ vocalization and /a&#618;/ monophthongization (e.g., Labov, Ash, &amp; Boberg, 2004). Even when investigating inherently continuous variables such as vowel quality, research on social meanings also tends to bin variables into discrete categories (<xref ref-type="bibr" rid="B52">Villarreal, 2018</xref>). In the present study, we address these gaps by investigating whether listeners&#8217; judgments of aspects of intonation are sensitive to the strength of the variable of interest in the phonetic signal.</p>
<p>We pursued these questions about intonational variation and social meaning via a task in which listeners rated samples of President Barack Obama&#8217;s speech on the degree of &#8216;sounding black.&#8217;<xref ref-type="fn" rid="n1">1</xref> Critical stimuli contained either one or more L+H* pitch accents or no L+H* pitch accents. The L+H* pitch accent has been shown in production studies to be a resource for performance of African American identity (<xref ref-type="bibr" rid="B15">Holliday, 2016</xref>; <xref ref-type="bibr" rid="B29">McLarty, 2018</xref>). Pitch accents in critical stimuli also varied according to degree of phonetic extremeness (i.e., the magnitude of F0 excursions). Listeners perceived stimuli with at least one L+H* token as sounding more black than those without, but only for stimuli with more phonetically extreme L+H* realizations (i.e., those with a larger difference between F0 maximum and minimum). These findings contribute to our understanding of how listeners make ethnic judgments based on intonational variation, and how listeners assign social meaning to gradient phonetic variation.<xref ref-type="fn" rid="n2">2</xref></p>
<sec>
<title>1.1. Ethnic identification in the U.S.</title>
<p>A body of linguistic research on ethnic identification dating back nearly 70 years has found that U.S. listeners are generally rather accurate (70&#8211;100%) at distinguishing black speakers from white speakers (cf. <xref ref-type="bibr" rid="B48">Thomas &amp; Reaser, 2004</xref>). Recent studies have attempted to unpack the role of suprasegmentals in ethnic identification. Thomas and Reaser (<xref ref-type="bibr" rid="B48">2004</xref>) found that listeners were equally accurate at ethnic identification for monotonized and unmodified stimuli, suggesting that listeners do not rely solely on F0 cues in ethnic identification. They also discovered that some cues relevant to pitch accents are recoverable even from monotonized stimuli (i.e., amplitude, duration, and segmental qualities), so it is conceivable that pitch accents may aid identification even in monotonized stimuli. Holliday and Jaggers (<xref ref-type="bibr" rid="B16">2015</xref>) examined listeners&#8217; ability to identify the ethnicity of U.S. politicians based on single-word stimuli, in order to assess the effects of voice quality on listener judgments. Building on some of the earlier findings of Purnell et al. (<xref ref-type="bibr" rid="B39">1999</xref>), Holliday and Jaggers found that several suprasegmental variables, including jitter and harmonics-to-noise ratio, influenced ethnic identification, though they note that a combination of multiple speakers and contexts may cause challenges in isolating speaker-specific variables influencing ethnic identification. For this reason, in the present study, we attempt to control for the effect of speaker-specific voice quality variation and more carefully isolate the prosodic variables that may affect ethnic identification by employing stimuli from a single speaker.</p>
</sec>
<sec>
<title>1.2. Intonational variation: Pitch accents</title>
<p>This study focuses on one particular type of intonational variable as a starting point for understanding how listeners may react to ethnically-linked suprasegmental features, using methods based in the auto-segmental/metrical (AM) intonational framework (<xref ref-type="bibr" rid="B34">Pierrehumbert, 1980</xref>). Essential to the AM theory is the idea that movements in fundamental frequency (F0), the main correlate of what we perceive as pitch, result from an underlying sequence of tones that determine their structure. In the AM theory, these tones are either low or high, and all movements of the pitch contour are composed of a series of low and high sequences. The labeling system for intonational phenomena that is based on the AM theory is called the Tones and Breaks Index system (ToBI). Each language, and indeed a number of dialects and varieties, have distinct ToBI systems that reflect the variety&#8217;s intonational specifications (<xref ref-type="bibr" rid="B2">Beckman &amp; Ayers-Elam, 1997</xref>). The ToBI system for Mainstream American English (MAE), originally developed by Beckman and Ayers-Elam (<xref ref-type="bibr" rid="B2">1997</xref>) and based on the findings of Pierrehumbert (<xref ref-type="bibr" rid="B34">1980</xref>), is the only ToBI system generally in use for examining variation within American English. MAE-ToBI has previously been used for descriptions of Jewish English (<xref ref-type="bibr" rid="B4">Burdin, 2015</xref>), Appalachian English (<xref ref-type="bibr" rid="B41">Reed, 2016</xref>), as well as African American Language (AAL) (<xref ref-type="bibr" rid="B15">Holliday, 2016</xref>; <xref ref-type="bibr" rid="B18">Jun &amp; Foreman, 1996</xref>; <xref ref-type="bibr" rid="B29">McLarty, 2018</xref>).<xref ref-type="fn" rid="n3">3</xref></p>
<p>MAE-ToBI contains two types of pitch movements: pitch accents, which occur on some stressed syllables, and edge tones, which occur at phrase boundaries. The current study focuses only on the movement of pitch accents, though it is important to note that we also tested for the perceptual effects of edge tones. This study focuses on the difference between two types of pitch accents in MAE: a simple high tone, labeled as H*, and a fall-rise, labeled as L+H*. Though other types of pitch accents exist, H* and L+H* are by far the most common pitch accents in most varieties of U.S. English, including AAL (<xref ref-type="bibr" rid="B5">Burdin et al., 2018</xref>).</p>
<p>Earlier studies have shown that pitch accents are perceptually salient for listeners and that na&#239;ve listeners can be trained to identify them quickly (<xref ref-type="bibr" rid="B30">McLarty, Vaughn, &amp; Kendall, 2017</xref>; <xref ref-type="bibr" rid="B46">Thomas, 2011</xref>). Especially relevant to the current study, studies such as Loman (<xref ref-type="bibr" rid="B28">1975</xref>), Holliday (<xref ref-type="bibr" rid="B15">2016</xref>), and McLarty (<xref ref-type="bibr" rid="B29">2018</xref>) have found that MAE and AAL exhibit different rates and contexts of use for H* versus L+H*. In particular, Loman (<xref ref-type="bibr" rid="B28">1975</xref>) and McLarty (<xref ref-type="bibr" rid="B29">2018</xref>) each found that L+H* pitch accents are more common in some varieties of AAL.</p>
<p>Recent work by Holliday (<xref ref-type="bibr" rid="B15">2016</xref>), Burdin (<xref ref-type="bibr" rid="B4">2015</xref>), and Reed (<xref ref-type="bibr" rid="B41">2016</xref>) <italic>inter alia</italic> has also found that a greater rate of use of the L+H* pitch accent may also be a resource in production for performance of different types of ethnic identity. For example, Holliday (<xref ref-type="bibr" rid="B15">2016</xref>) recorded 25 men (age 18&#8211;32) with one black parent and one white parent in Washington, DC to examine their rates of use of different types of pitch accents in ethnic identity performance. The participants were recorded in casual peer dyad conversations, and the analysis of their intonational patterns was taken from these recordings. A sociolinguistic interview also elicited ideologies about race and self-identifications. The participants who identified more as black, as opposed to multiracial or mixed, were more likely to use a greater quantity of L+H* accents than H* accents. This finding supports Loman&#8217;s (<xref ref-type="bibr" rid="B28">1975</xref>) and McLarty&#8217;s (<xref ref-type="bibr" rid="B29">2018</xref>) findings that L+H* is more prevalent in AAL than in MAE; also relevant for the current study, this finding demonstrates that speakers&#8217; production of intonational variation is gradient in terms of frequency.</p>
</sec>
<sec>
<title>1.3. Incrementality in intonation and perception</title>
<p>This study&#8217;s focus on intonational and suprasegmental variation presents an opportunity to address questions about phonetic detail and social meaning. One of the most significant recent advances in sociolinguistic theory has been the advent of sociophonetics (e.g., <xref ref-type="bibr" rid="B10">Foulkes &amp; Docherty, 2006</xref>), with the notion that paying attention to phonetic detail can enrich our understanding of sociolinguistic variation&#8212;especially for variables that have traditionally been considered binary or categorical (<xref ref-type="bibr" rid="B24">Labov, Ash, &amp; Boberg, 2006</xref>).</p>
<p>Although the binary treatment of phonetic variables reveals structure in sociolinguistic variation, a sociophonetically informed approach recognizes that the distribution of these variables&#8217; continuous acoustic correlates is not always compatible with discrete categorization. For example, Jacewicz and Fox (<xref ref-type="bibr" rid="B17">2018</xref>) use a continuous measure of /a&#618;/ monophthongization (trajectory length) to analyze preadolescent Appalachian English speakers. They find that these preadolescents produce variants that are more diphthongal than Appalachian adults but less diphthongal than central Ohio adults. The authors&#8217; continuous approach pays off, in other words, by revealing finer-grained phonetic variation than is suggested by the monophthong/diphthong binary.</p>
<p>At the same time as research on production in sociolinguistics has increasingly turned to phonetic detail, the role of such detail remains under-theorized and under-investigated in the study of social meaning. To that end, Podesva (<xref ref-type="bibr" rid="B38">2011</xref>) proposes a framework for salience in sociolinguistic variation that reconciles the roles of frequency and phonetic detail. He hypothesizes that salience takes one of two linguistic forms: &#8216;categorial salience&#8217; (frequent productions of a marked feature are salient) and &#8216;phonetic salience&#8217; (more extreme productions are salient). In particular, with respect to phonetic salience, Podesva argues that a more extreme production signals a stronger social meaning: &#8220;If an axis of phonetic variation indexes a particular social meaning, then outliers on that axis can be understood as the <italic>strongest indicators of meaning</italic>&#8221; (pp. 254, emphasis added).</p>
<p>These predictions about categorial and phonetic salience have been supported by a handful of findings on the distribution and social meaning of intonational variation in production. For example, Podesva (<xref ref-type="bibr" rid="B38">2011</xref>) found that one speaker constructed a &#8216;life of the party&#8217; persona by using acoustically extreme falling contours to imbue partying-related narrative elements with extra emphasis. Burdin et al.&#8217;s (<xref ref-type="bibr" rid="B5">2018</xref>) comparison of L+H* pitch accents in Jewish English, AAL, and Appalachian English showed that both categorical and continuous properties of pitch accents are sites for sociolinguistic differentiation. The authors found that, across communities, L+H* pitch accents differed in both rates of use and acoustic properties (e.g., peak F0, peak offset).</p>
<p>As far as we are aware, only a handful of perceptual studies have investigated how social meanings are affected by phonetic detail. Plichta and Preston (<xref ref-type="bibr" rid="B37">2005</xref>) presented U.S. listeners with a synthesized continuum from monophthongal to diphthongal /a&#618;/ and asked listeners to identify the speaker&#8217;s geographic origin along an axis running from the U.S. north to the U.S. south. Listeners not only associated monophthongal /a&#618;/ with the south and diphthongal /a&#618;/ with the north, they also placed successive continuum steps linearly along the north&#8211;south axis. D&#8217;Onofrio (<xref ref-type="bibr" rid="B9">2018</xref>) found that labeling a speaker as a &#8216;Business Professional&#8217; or &#8216;Valley Girl&#8217; cued U.S. listeners to classify more ambiguous [&#230;~&#593;] tokens as /&#230;/, with &#8216;Valley Girl&#8217; being especially associated with backer /&#230;/, compared to a &#8216;Chicago Bears Fan&#8217; label or no label at all. In an experiment with Californian listeners, Villarreal (<xref ref-type="bibr" rid="B51">2016</xref>) found significant correlations between speakers&#8217; raising of /&#230;/ in <italic>bad</italic> and <italic>glass</italic> and listeners&#8217; ratings on the scales &#8216;accented,&#8217; &#8216;doesn&#8217;t speak like me,&#8217; &#8216;unfamiliar,&#8217; and &#8216;not Californian.&#8217; Foulkes, Docherty, Khattab, and Yaeger-Dror (<xref ref-type="bibr" rid="B11">2010</xref>) found that listeners&#8217; identification of Tyneside children&#8217;s gender was affected by two continuous measures (amplitude and F0) as well as several categorical measures; however, the authors also report significant correlations between amplitude and F0 in stimuli, suggesting potential issues with collinearity in the modeling procedure. In terms of voice quality, Szakay (<xref ref-type="bibr" rid="B44">2012</xref>) found that in New Zealand, ethnic identification was affected by several continuous voice quality measures; speakers with higher mean H1&#8211;H2 (a measure of creakiness) were likelier to be identified as M&#257;ori.</p>
<p>The present study seeks to expand our understanding of the relationship between phonetic detail and social meaning by investigating this relationship through the lens of intonational variation. Building on Podesva (<xref ref-type="bibr" rid="B38">2011</xref>), we hypothesize that the social meanings of continuous variables will exhibit what we call <italic>incrementality</italic>: a monotonic relationship between the variable&#8217;s phonetic extremeness and the strength of the social meaning it elicits in perceivers.<xref ref-type="fn" rid="n4">4</xref> We focus on pitch accents, which are ideally suited to this question as they vary both in category (e.g., H* versus L+H*) and phonetic shape (e.g., peak offset, rise slope).</p>
</sec>
</sec>
<sec sec-type="methods">
<title>2. Methods</title>
<p>This study was designed to address three central research questions:</p>
<list list-type="order">
<list-item><p>How do pitch accents affect listener judgments of ethnic identity? In particular, does the L+H* pitch accent carry a social meaning of blackness in perception, as it does in production?</p></list-item>
<list-item><p>To what extent are the ethnicity-based social meanings of these pitch accents mediated by incremental phonetic differences?</p></list-item>
<list-item><p>What other aspects of voice quality affect listener judgments of ethnicity?</p></list-item>
</list>
<p>These questions were investigated via a perceptual task in which listeners rated 120 samples of President Barack Obama&#8217;s speech with respect to how much they thought he &#8216;sounded black&#8217; in each particular sample.</p>
<sec>
<title>2.1. Open-guise versus matched-guise technique</title>
<p>This task used the &#8216;open-guise technique&#8217; (OGT) (<xref ref-type="bibr" rid="B43">Soukup, 2013</xref>); as in the more common matched-guise technique (MGT), OGTs offer insight into the social meanings of a focal feature, variety, or language, by comparing listeners&#8217; reactions to stimuli differing only by the focal linguistic structure (e.g., <xref ref-type="bibr" rid="B6">Campbell-Kibler, 2009</xref>). Unlike the OGT, the MGT axiomatically hinges on listeners&#8217; belief that they are listening to different speakers (<xref ref-type="bibr" rid="B12">Giles &amp; Billings, 2004</xref>; <xref ref-type="bibr" rid="B39">Purnell et al., 1999</xref>); otherwise, it is assumed that listeners will not differentiate guises on personal characteristics that are considered intrapersonally stable qualities (e.g., intelligence). In OGTs, by contrast, listeners are openly informed that they are hearing the same speaker in different guises. Soukup (<xref ref-type="bibr" rid="B43">2013</xref>) shows that listeners responded differently to standard versus dialectal Austrian German guises in both OGT and MGT settings (with the OGT actually yielding stronger effects for some scales), undermining MGTs&#8217; key assumption about different speakers.</p>
<p>In the present study, we assumed that listeners (all from the United States) were highly likely to recognize our stimulus speaker, President Barack Obama, necessitating an OGT rather than MGT approach. We openly informed our listeners, &#8220;This study is designed to test how people respond to different speech excerpts from the same speaker.&#8221; In so doing, we rejected the type of instrumental task framing often used in MGTs, such as evaluating prospective radio newsreaders (<xref ref-type="bibr" rid="B25">Labov et al., 2011</xref>; <xref ref-type="bibr" rid="B52">Villarreal, 2018</xref>). By contrast, Obama represented an ideal stimulus speaker to test our hypotheses, as his ability to command both AAL and MAE is well-known by the general public (<xref ref-type="bibr" rid="B1">Alim &amp; Smitherman, 2012</xref>); our use of the OGT took advantage of this awareness. In the discussion, we make recommendations about the appropriateness of OGT versus MGT.</p>
</sec>
<sec>
<title>2.2. Stimulus creation</title>
<p>The 120 stimuli were based on excerpts of President Barack Obama&#8217;s spontaneous speech from two different 2016 television interviews with Gayle King, a black broadcast journalist who co-anchors the <italic>CBS This Morning</italic> news program (<xref ref-type="bibr" rid="B19">Kaplan, 2016</xref>). Each stimulus excerpt was based on a single Intonational Phrase (IP) unit, ranging from 0.4 to 2.3 seconds in duration (median 0.9 seconds). Following Pierrehumbert and Hirschberg (<xref ref-type="bibr" rid="B36">1990</xref>) as well as subsequent works utilizing their methods, we identified IPs through looking for pausing and phrase-final lengthening, as well as the presence of characteristic boundary tones and smaller intermediate phrase units contained within the IPs. We attempted to select short phrases that were fairly semantically bland to avoid overly tilting responses in one direction, though it is impossible to completely control for content in listening tasks.</p>
<p>Sixty excerpts were selected: 20 critical excerpts and 40 filler excerpts. Ten critical excerpts were &#8216;H* phrases,&#8217; which contained between 1&#8211;3 H* pitch accents and 0 L+H* accents; and ten were &#8216;L+H* phrases,&#8217; which contained between 1&#8211;3 L+H* pitch accents and 0&#8211;2 H* accents. This imbalanced definition of H* versus L+H* phrases was necessary since L+H* pitch accents are relatively rarer, even in AAL (<xref ref-type="bibr" rid="B5">Burdin, Holliday, &amp; Reed, 2018</xref>), so it was not possible to find enough excerpts that contained only L+H* accents. Filler excerpts contained 1&#8211;3 H* pitch accents and 0 L+H* accents.</p>
<p>In choosing excerpts, we intentionally sacrificed a degree of experimental control for the sake of presenting listeners with natural, spontaneously produced stimuli rather than unnatural, lab-like speech. The benefit of using spontaneous stimuli is that it more closely models real-world perception conditions, as listeners perceive spontaneous and read speech (including oratory) differently (<xref ref-type="bibr" rid="B6">Campbell-Kibler, 2009</xref>; <xref ref-type="bibr" rid="B16">Holliday &amp; Jaggers, 2015</xref>). The drawback is that the distribution of H* versus L+H* pitch accents across stimuli prevented us from addressing Podesva&#8217;s (<xref ref-type="bibr" rid="B38">2011</xref>) hypothesis about categorial salience; at the same time, total experimental control over stimuli is impossible to obtain, as features co-occurring in the stimuli can always shape interpretation of features of interest (<xref ref-type="bibr" rid="B26">Leach, Watson, &amp; Gnevsheva, 2016</xref>), including propositional content (<xref ref-type="bibr" rid="B6">Campbell-Kibler, 2009</xref>). (We explore this issue further in a post hoc analysis of L+H* phrases.)</p>
<p>The critical stimuli were created by manipulating critical excerpts to four manipulation steps, with the original excerpt as Step 1. Steps 2, 3, and 4 were created by making pitch accents&#8217; F0 minima and maxima successively more extreme. With each manipulation step, H* and L+H* maxima were increased by a semitone, and L+H* minima were decreased by a half-semitone. For example, the H* pitch accent in the top panel of Figure <xref ref-type="fig" rid="F1">1</xref> has an F0 maximum at 118.3 Hz in step 2 and 125.2 Hz in step 3, a one-semitone difference; the L+H* pitch accent in the bottom panel has an F0 minimum at 101.6 Hz in step 1 and 99.1 Hz, a half-semitone difference. (In some cases, it was not possible to make the manipulations exactly one or one-half semitone.) We based F0 manipulations on semitones rather than constant magnitudes because semitones are psychoacoustically comparable regardless of the pitch accent&#8217;s initial F0 (e.g., the difference between 100 and 105 Hz sounds much larger than the difference between 200 and 205 Hz). The first author created stimuli by hand using the Manipulation utility in Praat (<xref ref-type="bibr" rid="B3">Boersma &amp; Weenink, 2015</xref>). Both authors listened to all manipulated critical stimuli and confirmed that they sounded natural.</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Original (Step 1) and manipulated (Steps 2&#8211;4) versions of pitch accents in stimuli: H* pitch accent in <italic>would</italic> (top) and L+H* pitch accent in <italic>all</italic> (bottom).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6261/file/75600/"/>
</fig>
<p>Filler stimuli were created by modifying the final syllable of filler excerpts to include percepts of creaky voice: low F0 and damped pulses (<xref ref-type="bibr" rid="B20">Keating, Garellek, &amp; Kreiman, 2015</xref>). A Praat script modified alternating cycles of the final syllable by lengthening their duration and lowering their amplitude. As with critical stimuli, both authors listened to all manipulated filler stimuli and confirmed that they sounded natural.</p>
</sec>
<sec>
<title>2.3. Task design</title>
<p>The task was administered via an online survey hosted by Qualtrics. In each of 120 randomly ordered trials, listeners heard a single stimulus auto-play twice and responded to the question &#8220;How black or white does Obama sound here?&#8221; on a continuous unit-less slider bar with &#8220;very black&#8221; and &#8220;very white&#8221; on opposite poles. As the recognizability of President Obama&#8217;s voice would have likely rendered ineffective the type of instrumental task framing often used in MGTs (e.g., rating prospective radio newsreaders, as in <xref ref-type="bibr" rid="B25">Labov et al., 2011</xref>), we eschewed such framing; we instead informed listeners, &#8220;This study is designed to test how people respond to different speech excerpts from the same speaker.&#8221; Listeners then completed a demographic questionnaire and were invited to comment on the task (see Appendix A).</p>
<p>The survey was distributed via social network sampling in May 2017, with a raffle incentive for one randomly selected listener to win an <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://amazon.com">Amazon.com</ext-link> gift card. The listener sample for analysis contains 79 American English-speaking listeners who self-identified as black and/or white. Of these listeners, 24% self-identified as black and 77% as white (one listener identified as both); 65% identified as female and 35% as male. The majority of listeners also self-identified as politically liberal and indicated that they overwhelmingly approved of Obama&#8217;s presidency; in particular, on a 1&#8211;7 scale (where 7 indicated &#8220;very liberal&#8221; and &#8220;strongly approve of Obama&#8221;), the median rating was 6 on both scales, and 91% of listeners rated 5 or above on <italic>both</italic> scales. In this respect, the listener sample is not representative of the United States voting population; however, our intent was not to survey a sample spanning the political spectrum but rather to determine how some listeners judge ethnicity based on intonational and voice quality variation (we return to this point in the Discussion).</p>
<p>As mentioned above, both authors listened to all stimuli and confirmed that they sounded natural. As a further check on stimulus naturalness, we coded listeners&#8217; responses to the final two questionnaire items: &#8220;How did the clips sound to you?&#8221; and &#8220;Do you have any other comments on the clips or on the survey?&#8221; Based on listeners&#8217; responses to these questions, the second author developed eight true-or-false codes that described sentiments listeners expressed in their responses and coded responses accordingly (with a single response capable of being coded &#8220;true&#8221; in multiple categories). For example, 21% of listeners reported something amiss with the quality of the clips (although numerous listeners commented positively about the clips&#8217; quality). More information about these codes, including examples, can be found in Appendix B. As we discuss below, however, none of these codes significantly improved our model of intonation results, so we did not find evidence that they impacted listeners&#8217; perceptions of the speaker&#8217;s blackness.</p>
<p>Slider-bar positions were converted to real numbers between 0 (&#8220;very white&#8221;) and 100 (&#8220;very black&#8221;) and standardized by listener to control for variable usage of the continuous slider bar. All results are reported in unit-less standard deviations (i.e., z-scores); the average listener&#8217;s standard deviation was 16.6, so a difference of 1 standard deviation can be interpreted as a difference of roughly one-sixth of the length of the slider bar for the average listener.</p>
<p>Our task was specifically designed to address the first two research questions, about the role of pitch accents and phonetic incrementality in affecting listener judgments of ethnicity; we first present the analysis of intonation features. We then describe a post hoc analysis of voice quality characteristics that addressed the third research question, about the role of other voice quality features in affecting listener judgments of ethnicity.</p>
</sec>
</sec>
<sec>
<title>3. Intonation analysis</title>
<p>We compared linear mixed-effects models of standardized ratings to find the predictor structure that best modeled the data in critical trials, via the lmerTest package for R (<xref ref-type="bibr" rid="B23">Kuznetsova, Brockhoff, &amp; Christensen, 2016</xref>; <xref ref-type="bibr" rid="B40">R Core Team, 2018</xref>). The predictors that we tested were phrase type (H* versus L+H*), manipulation step, edge tone, nuclear pitch accent, stimulus duration, and numerous listener effects (race, gender, political ideology, approval of Obama&#8217;s presidency, education, use of desktop versus mobile to complete survey, hometown, geographic mobility, experience with linguistics, musical experience, and qualitative questionnaire codes). Unfortunately, the distribution of H* and L+H* tokens in stimuli precluded predictors for the number of H* and number of L+H* pitch accents in critical trials. We also included random intercepts for excerpts as nested within phrase type, as each excerpt exclusively belonged to one of the two phrase types. Since ratings were standardized by listener, by-listener random intercepts would be redundant.</p>
<p>Table <xref ref-type="table" rid="T1">1</xref> presents a summary of the best model for listener rates of blackness, which included predictors of phrase type, manipulation step, and their interactions.</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>Summary of best model of listener ratings of blackness. Degrees of freedom estimated via Satterthwaite approximations (<xref ref-type="bibr" rid="B42">Satterthwaite, 1946</xref>). Significance: * <italic>p</italic> &lt; 0.05.</p>
</caption>
<table>
<tr>
<th align="left" valign="top"></th>
<th align="center" valign="top">Estimate</th>
<th align="center" valign="top"><italic>SE</italic></th>
<th align="center" valign="top"><italic>d.f.</italic></th>
<th align="center" valign="top"><italic>t</italic></th>
<th align="center" valign="top"><italic>p</italic></th>
</tr>
<tr>
<td colspan="6"><hr/></td>
</tr>
<tr>
<td align="left" valign="top">(Intercept)</td>
<td align="right" valign="top">&#8211;0.0409</td>
<td align="right" valign="top">0.1065</td>
<td align="right" valign="top">23.1</td>
<td align="right" valign="top">&#8211;0.384</td>
<td align="right" valign="top">0.7045&#160;&#160;</td>
</tr>
<tr>
<td align="left" valign="top">PhrTypeL+H*</td>
<td align="right" valign="top">0.0172</td>
<td align="right" valign="top">0.1507</td>
<td align="right" valign="top">23.1</td>
<td align="right" valign="top">0.114</td>
<td align="right" valign="top">0.9103&#160;&#160;</td>
</tr>
<tr>
<td align="left" valign="top">Step2</td>
<td align="right" valign="top">0.0041</td>
<td align="right" valign="top">0.0458</td>
<td align="right" valign="top">6056</td>
<td align="right" valign="top">0.089</td>
<td align="right" valign="top">0.929&#160;&#160;</td>
</tr>
<tr>
<td align="left" valign="top">Step3</td>
<td align="right" valign="top">&#8211;0.0214</td>
<td align="right" valign="top">0.0459</td>
<td align="right" valign="top">6056</td>
<td align="right" valign="top">&#8211;0.466</td>
<td align="right" valign="top">0.6415&#160;&#160;</td>
</tr>
<tr>
<td align="left" valign="top">Step4</td>
<td align="right" valign="top">0.0132</td>
<td align="right" valign="top">0.0459</td>
<td align="right" valign="top">6056</td>
<td align="right" valign="top">0.287</td>
<td align="right" valign="top">0.7737&#160;&#160;</td>
</tr>
<tr>
<td align="left" valign="top">PhrTypeL+H*:Step2</td>
<td align="right" valign="top">0.0297</td>
<td align="right" valign="top">0.065</td>
<td align="right" valign="top">6056</td>
<td align="right" valign="top">0.457</td>
<td align="right" valign="top">0.6475&#160;&#160;</td>
</tr>
<tr>
<td align="left" valign="top">PhrTypeL+H*:Step3</td>
<td align="right" valign="top">0.1314</td>
<td align="right" valign="top">0.065</td>
<td align="right" valign="top">6056</td>
<td align="right" valign="top">2.02</td>
<td align="right" valign="top">0.0434*</td>
</tr>
<tr>
<td align="left" valign="top">PhrTypeL+H*:Step4</td>
<td align="right" valign="top">0.1047</td>
<td align="right" valign="top">0.065</td>
<td align="right" valign="top">6056</td>
<td align="right" valign="top">1.609</td>
<td align="right" valign="top">0.1076&#160;&#160;</td>
</tr>
</table>
</table-wrap>
<p>As is evident from this model, listener ratings of blackness tended to increase with the more extreme step manipulations, though this is only statistically significant for L+H* phrases. Also notable is that the model revealed no significant listener effects for gender, race, region, education, or political affiliation, indicating that listeners were remarkably similar in their ratings regardless of a number of potentially influential demographic factors. While previous studies have generally found that in-group community members may perform better in ethnic identification tasks (cf. <xref ref-type="bibr" rid="B48">Thomas &amp; Reaser, 2004</xref>), there were no such effects observed here. In addition, none of the qualitative questionnaire codes significantly improved the model; this means that, for example, although some listeners commented negatively on the quality of the stimuli, we have no evidence that whether or not listeners commented on stimulus quality affected listener perceptions of blackness.</p>
<p>These results must be interpreted with caution, however, in light of their small effect size. The sole significant term in Table <xref ref-type="table" rid="T1">1</xref>, PhrTypeL+H*:Step3 (which has a <italic>p</italic> value just under the predetermined &#945; level of 0.05), differs from the intercept by about 0.17 standard deviations, or about 3 &#8216;notches&#8217; on the 0&#8211;100 slider bar (with the average listener&#8217;s standard deviation being 16.6). Indeed, an R<sup>2</sup> calculation using the R package piecewiseSEM (<xref ref-type="bibr" rid="B27">Lefcheck, 2016</xref>) revealed that the model&#8217;s fixed-effects predictor structure accounted for less than 1% of the variance in ratings, while random effects&#8212;the effect of individual excerpts&#8212;accounted for 11.3% of the variance.<xref ref-type="fn" rid="n5">5</xref> With that caveat in mind, we proceed to discuss what these results mean.</p>
<sec>
<title>3.1. Results by phrase type</title>
<p>The model indicated no main effect of phrase type on listener ratings of blackness, indicating that pitch accent alone did not trigger different blackness ratings. Figure <xref ref-type="fig" rid="F2">2</xref> shows this result, with results for H* stimuli in the left panel and L+H* stimuli in the right panel. As is evident in this figure, listener ratings of blackness were remarkably similar for the H* phrases at each step, though the L+H* phrases showed greater differences between manipulation step. There was also no main effect of manipulation step on listener ratings of blackness.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Fitted model predictions for listener ratings of blackness by phrase type and manipulation step. Error bars represent 95% confidence intervals.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6261/file/75601/"/>
</fig>
</sec>
<sec>
<title>3.2. Results by manipulation step</title>
<p>Though the main effect of phrase type failed to reach significance, the model indicated a significant interaction between phrase type and manipulation step, with more extreme L+H* phrases rated as sounding blacker than less extreme L+H* phrases, and no perceived blackness difference for H* phrases regardless of step. Figure <xref ref-type="fig" rid="F3">3</xref> presents these results, with each panel representing a manipulation step. This figure indicates that listeners appear to interpret the more phonetically extreme L+H* realizations (greater difference between F0 minimum and F0 maximum within a L+H* pitch accent) as blacker, but this is not the case for the more extreme H* realizations (which only had higher F0 maxima).</p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Fitted model predictions for listener ratings of blackness by manipulation step and phrase type. Error bars represent 95% confidence intervals.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6261/file/75602/"/>
</fig>
<p>This model also implies that listener judgments of blackness are affected by more than just pitch accent type and phonetic shape. As mentioned above, these results must be interpreted with caution, especially in light of the fact that the model&#8217;s fixed-effects predictor structure accounted for less than 1% of the variance in ratings, while random effects&#8212;the effect of individual excerpts&#8212;accounted for 11.3% of the variance. In other words, listeners were much more attuned to features varying by excerpt, such as segmental, semantic, pragmatic, or voice quality characteristics, than the type and phonetic shape of pitch accents. However, this small effect size may represent an inherent challenge to studies of prosody, since the highly nested nature of such variables causes them to be difficult to isolate from one another. Despite this challenge, the finding of a significant difference here may be a step in the direction of discovering how these variables may operate both independently and together. The small effect size of this intonation effect motivated the post hoc analysis of voice quality features.</p>
</sec>
</sec>
<sec>
<title>4. Voice quality analysis</title>
<p>Our perceptual experiment was specifically designed to test predictions about how listener judgments of ethnicity are influenced by the type and phonetic shape of pitch accents; however, sociophoneticians have long suspected that voice quality characteristics may also influence listener judgments of ethnicity (e.g., <xref ref-type="bibr" rid="B16">Holliday &amp; Jaggers, 2015</xref>; <xref ref-type="bibr" rid="B39">Purnell et al., 1999</xref>). In line with Purnell et al. (<xref ref-type="bibr" rid="B39">1999</xref>), we conducted a post hoc analysis of perceived blackness ratings to determine if and how a number of voice quality measures were influential in shaping listener judgments. In particular, the results of their study indicate dialect-level differences in both harmonics to noise ratio (HNR) and peak pitch ratio, so we hypothesized that these same variables may also be of interest in the current study.</p>
<p>We ran a Praat script on critical stimuli to extract several measures that, according to previous studies, may pattern differently in AAL versus MAE: phrase speech rate, pitch ratio (<xref ref-type="bibr" rid="B16">Holliday &amp; Jaggers, 2015</xref>), peak delay (<xref ref-type="bibr" rid="B15">Holliday, 2016</xref>; <xref ref-type="bibr" rid="B41">Reed, 2016</xref>), jitter (<xref ref-type="bibr" rid="B16">Holliday &amp; Jaggers, 2015</xref>), shimmer (ibid.), HNR (<xref ref-type="bibr" rid="B39">Purnell et al., 1999</xref>), and intensity average (ibid.).<xref ref-type="fn" rid="n6">6</xref> Phrase speech rate was calculated as the stimulus&#8217;s duration divided by the number of syllables. Pitch ratio was calculated as the stimulus&#8217;s maximum F0 (in Hz) divided by its minimum F0. The remaining measures were calculated for each pitch accent in each stimulus; since listeners reacted not to individual PAs but whole stimuli, for stimuli with multiple PAs we treated the mean of each PA&#8217;s measurement as the measurement for that stimulus (e.g., we defined the jitter measurement for a stimulus with three PAs as the mean of the PAs&#8217; jitter measurements). Peak delay was calculated as the time difference between nucleus onset and hand-annotated pitch accent time. Jitter (relative average perturbation), shimmer (local amplitude perturbation), HNR, and intensity average (mean dB) were all calculated for the nucleus. An F0 floor of 75 Hz was used for all relevant measures in order to avoid erroneous measurements of non-periodic speech; we otherwise used Praat&#8217;s default settings for all measurement functions.</p>
<p>As with the intonation analysis, we modeled standardized ratings via linear mixed-effects models. Because the intonation analysis revealed differences in patterning of responses to H* versus L+H* stimuli, we fit separate models to H* versus L+H* critical trials. We included manipulation step in these models to determine whether the intonation analysis&#8217;s findings about the role of manipulation step&#8212;significantly affecting listener ratings of blackness in L+H* stimuli but not H* stimuli&#8212;remained after considering voice quality features. These models also included random intercepts for excerpts and random by-excerpt slopes for the manipulation step factor. Voice quality measures were normalized (z-scored) to account for widely differing measurement scales.</p>
<p>To account for likely collinearity of voice quality measures (e.g., jitter and pitch ratio are all different measures of changes in fundamental frequency), we adopted a model-comparison strategy that iteratively added interaction terms to the models based on correlations between measures. We first ran baseline models that included all voice quality measures as main effect predictors with no interactions. (Again, these models also included random intercepts for excerpts and random by-excerpt slopes for the manipulation step factor.) We then checked these baseline models for correlations between voice quality measures; any correlations with an absolute value correlation coefficient greater than 0.4 in either model were added as interaction terms into both models. After running these models, we again added interaction terms (including three-way interactions) based on correlations between voice quality terms. The resulting models included the following interactions: phrase speech rate &#215; peak delay &#215; HNR, shimmer &#215; jitter &#215; HNR, pitch ratio &#215; intensity average. For both the H* and L+H* models, each successive model represented a significant improvement in model fit at an &#945; = .05 significance threshold.</p>
<sec>
<title>4.1. Voice quality results</title>
<p>Summaries of fixed effects for the voice quality models are in Appendix C. The voice quality model for H* critical trials revealed that few voice quality measures significantly affected listener perceptions of blackness: phrase speech rate and the interaction of peak delay and HNR. Phrase speech rate (seconds per syllable) had a positive effect on listener perceptions of blackness, with slower phrases rated blacker. While the model returned positive estimates for the effects of peak delay and HNR, neither of these main effects reached significance. Rather, the effect of peak delay on listener perceptions of blackness was constrained by the phrase&#8217;s HNR. As Figure <xref ref-type="fig" rid="F4">4</xref> shows, phrases with longer peak delay were rated blacker, but only if HNR was sufficiently high. This co-patterning of variables suggests that listeners may be attuning to a threshold of combined characteristics in order to make judgments, particularly in the absence of ethnolinguistically-salient intonational differences such as the L+H* pitch accent. Together, the roles of peak delay and phrase speech rate suggest a possible salience effect, as both relate to vowel duration; conceivably, longer vowels (which co-pattern with slower speech rates) with longer intonation rises can better carry indexes of social meaning.<xref ref-type="fn" rid="n7">7</xref></p>
<fig id="F4">
<label>Figure 4</label>
<caption>
<p>H* model predictions for perceived blackness ratings by peak delay (seconds) and HNR (dB). The five facets display peak delay slopes at the minimum, first quartile, median, third quartile, and maximum values for HNR among H* stimuli.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6261/file/75603/"/>
</fig>
<p>As with the H* model, few predictors reached significance in the L+H* model&#8212;including just one voice quality measure, jitter. Among L+H* stimuli, phrases with less jitter were rated blacker, suggesting that listeners are sensitive to the interaction of F0 movement and local periodic perturbations. Notably, the measures affecting listener perceptions of blackness did not overlap for H* versus L+H* phrases; the jitter term in the H* model, and the phrase speech rate &amp; peak delay &#215; HNR terms in the L+H* model, did not even approach significance. This finding provides additional evidence that listeners may respond to different intonation and voice quality cues in phrases containing L+H* pitch accents than those not containing L+H* pitch accents. As L+H* accents are far less common than H* accents, it is possible that L+H* accents cue listeners to adjust their expectations as to markers of ethnic identification.</p>
<p>In addition, manipulation step was significant in the L+H* voice quality model (manipulation steps 3 and 4 were rated blacker than steps 1 and 2) but not the H* voice quality model. This finding corroborates the generalization that the percept of blackness is subject to phonetic incrementality only with respect to the more socially marked L+H* pitch accent. However, this finding is tempered by the fact that the fixed effects in the L+H* model accounted for less than 3% of the variance, as compared with 10.2% for the fixed effects in the H* model (Table <xref ref-type="table" rid="T2">2</xref>); in other words, there remain properties of the stimuli with L+H* accents that listeners are reacting to, above and beyond the intonational and voice quality features that previous studies have suggested are implicated in differentiating AAL from MAE.</p>
<table-wrap id="T2">
<label>Table 2</label>
<caption>
<p>R<sup>2</sup> values (percentage of variance accounted for) for voice quality models, calculated via R package piecewiseSEM (<xref ref-type="bibr" rid="B27">Lefcheck, 2016</xref>).</p>
</caption>
<table>
<tr>
<th align="left" valign="top"></th>
<th align="center" valign="top">Fixed-effects R<sup>2</sup></th>
<th align="center" valign="top">Random-effects R<sup>2</sup></th>
<th align="center" valign="top">Total R<sup>2</sup></th>
</tr>
<tr>
<td colspan="4"><hr/></td>
</tr>
<tr>
<td align="left" valign="top">H* model</td>
<td align="right" valign="top">10.2%</td>
<td align="right" valign="top">4.3%</td>
<td align="right" valign="top">14.5%</td>
</tr>
<tr>
<td align="left" valign="top">L+H* model</td>
<td align="right" valign="top">2.7%</td>
<td align="right" valign="top">24.7%</td>
<td align="right" valign="top">27.3%</td>
</tr>
</table>
</table-wrap>
<p>In short, the voice quality analysis found that listeners relied on multiple acoustic cues&#8212;beyond those pertaining to pitch accents&#8217; type or phonetic shape&#8212;in making judgments of perceived blackness; crucially, in the presence of an L+H* pitch accent listeners not only relied on different voice quality cues than in the absence of one, but they apparently relied to a much greater degree on cues other than those relating to intonation or voice quality. This finding suggests a fundamental difference in how listeners judge phrases in the presence of an L+H* pitch accent, although this is an open question for future study. More broadly, this finding further supports the claim that understanding the interrelated nature of prosodic variables is a necessary part of their description.</p>
</sec>
</sec>
<sec>
<title>5. Discussion</title>
<p>To summarize, this study has demonstrated that listeners are sensitive to the details of phonetic realizations of the H* and L+H* pitch accents in declaratives, and that a larger difference between the F0 maximum and minimum within L+H* pitch accents appears to cause listeners to rate a speaker (in this case, President Barack Obama) as sounding blacker. However, the difference between H* and L+H* pitch accent phrases alone is not sufficient to trigger this judgment; it is the actual realization of the pitch accents themselves that listeners seem to attune to. In addition to pitch accent type and phonetic shape, listeners also attend to voice quality cues in judging blackness, though the relevant cues are different for H* versus L+H* phrases: speech rate, peak delay, and harmonics to noise ratio for H* phrases, jitter for L+H* phrases. There is also some evidence that the number of L+H* and H* pitch accents in a phrase also affect listener judgments of blackness. We also obtained an unexpected finding with respect to speech rate; among H* stimuli, slower phrases were perceived blacker than faster phrases, which could possibly indicate that speakers have different expectations related to ethnolinguistic variation and speech rate (<xref ref-type="bibr" rid="B21">Kendall, 2013</xref>) or that longer vowels provide a greater site for the apprehension of social meaning. In this section we discuss the implications of these findings in more depth.</p>
<sec>
<title>5.1. Intonation</title>
<p>This study&#8217;s results show that in a perception task, listeners appear to be sensitive not only to the phonological category of pitch accents, but also their phonetic realization, as listeners appear to be sensitive to increasingly extreme manipulations of F0 within a single pitch accent type. In the traditional AM model of intonational phonology, pitch accent and edge tones have largely been binned into discrete categories, with meaning presumed to be attached to those categories and their combinations (<xref ref-type="bibr" rid="B36">Pierrehumbert &amp; Hirschberg, 1990</xref>). The results presented here provide further motivation for considering intonational variation on a phonetic as well as a phonological level. This study also provides further motivation for the development of ethnolinguistic variety-specific ToBI models as well as phonetic methods for studying intonational variation cross-dialectically. While we have employed the MAE-ToBI conventions (<xref ref-type="bibr" rid="B2">Beckman &amp; Ayers-Elam, 1997</xref>) in this study, the nature of the intonational system of AAL has not yet been fully described (<xref ref-type="bibr" rid="B29">McLarty, 2018</xref>; <xref ref-type="bibr" rid="B47">Thomas, 2015</xref>). As the current study&#8217;s results provide evidence that listeners are sensitive to differences in the realization of F0 and timing of the L+H* pitch accent, future studies should examine whether the tonal inventory of AAL differs from MAE, as this could be one element that triggers the observed differences in listener judgments.</p>
<p>Relatedly, as much of the work on prosody has focused on the meaning of intonational contours in an imagined Standard American English as opposed to in specific varieties, it is clear that much more work is needed on both variation in speaker production and listener perception of contour meaning. Though the current study did not reveal differences in perception of &#8216;sounding black&#8217; conditioned by listener demographics, future work should explore how such perceptions could potentially be affected by listeners with different backgrounds and sociolinguistic experiences.</p>
<p>This point about the role of demographics is especially relevant because (as mentioned above) the listener sample was overwhelmingly liberal and approving of Obama&#8217;s presidency, more so than the US population at large. While this is not an issue for the present study&#8212;our aim was not to achieve political representativeness but rather to ascertain how intonational variation affected perceptions of blackness within a population of US listeners&#8212;it does contextualize the results. Theoretical frameworks that take as primary the role of experience in forming linguistic representations (e.g., Exemplar Theory, <xref ref-type="bibr" rid="B35">Pierrehumbert, 2016</xref>) would take the standpoint that listeners who are more inclined to listen to President Obama would have a greater opportunity to hear him in multiple situations, thus facilitating their awareness of his style-shifting; it is thus conceivable that the small-sized effects uncovered in this study would not reach significance in a sample more representative of the US political spectrum. This is a question open for future work to address.</p>
</sec>
<sec>
<title>5.2. Ethnic identification</title>
<p>In their 2004 study and summary of the body of research on ethnic identification of white and black speakers and the U.S., Thomas and Reaser reveal gaps in our knowledge about what triggers judgments of speakers as &#8216;black&#8217; or &#8216;white.&#8217; Most ethnic identification studies have focused on segmental features, at least in part due to the fact that so little is known about how non-standard varieties of American English employ intonational variation, though it is the case that such studies on prosodic variables have been carried out outside the U.S. (<xref ref-type="bibr" rid="B44">Szakay, 2012</xref>; <xref ref-type="bibr" rid="B49">Todd, 2002</xref>). While a number of segmental features, such as vowel quality, have been identified as important in triggering listener judgments, researchers still know relatively little about how suprasegmental features may contribute to these judgments. Dating back to the 1970s, researchers such as Tarone (<xref ref-type="bibr" rid="B45">1973</xref>) and Loman (<xref ref-type="bibr" rid="B28">1975</xref>) have suspected that suprasegmental features played a serious role in triggering these judgments, though few studies have been able to isolate the specific intonational and suprasegmental features involved. The results of the current study, especially those related to the fact that speakers are able to provide consistent judgments of how a speaker whose race is known to them adheres to their ideologies about what it means to &#8216;sound black,&#8217; provide evidence that it may be possible to isolate the variables of interest using a single-speaker model. This has the advantage of eliminating other types of variation that are inherent in studies with multiple speakers, for whom it is impossible to control every level of linguistic variation, which may be important especially in light of our findings on the effects of voice quality.</p>
<p>This study also builds on the findings of Purnell et al. (<xref ref-type="bibr" rid="B39">1999</xref>) as well as Thomas and Reaser (<xref ref-type="bibr" rid="B48">2004</xref>), and Holliday and Jaggers (<xref ref-type="bibr" rid="B16">2015</xref>) by providing further evidence that a number of voice quality features, including jitter, HNR, and speech rate may be involved in triggering ethnicity judgments. The pattern that we observed wherein there appear to be important interactions of intonational and voice quality features obviates the need for more controlled studies that simultaneously focus on a number of suprasegmental features. Listeners appear not only to be sensitive to both intonational and voice quality features but also the ways in which they combine to create sociolinguistic meaning.</p>
<p>It is worth reiterating here the small effect size that we found in our intonation model, in which fixed effects accounted for just 1% of the variance in listener ratings of blackness. Some readers may interpret this small effect size and the proximity of the sole significant intonational model term&#8217;s <italic>p</italic> value (0.0434) to our predetermined &#945; level (0.05) as casting doubt upon the generality of the result. Although this effect size is modest, it is not without precedent in studies of sociolinguistic perception. Clopper (<xref ref-type="bibr" rid="B7">2010, p. 212</xref>), describing Clopper and Pisoni&#8217;s (<xref ref-type="bibr" rid="B8">2007</xref>) study of free classification of regional dialects of American English, notes &#8220;grouping accuracy was still rather poor overall, which may indicate attention to talker-specific differences instead of dialect-specific variation.&#8221; This greater attention to talker-specific differences parallels our finding that random effects accounted for a much greater percentage of the intonation model&#8217;s variance. Likewise, Villarreal (<xref ref-type="bibr" rid="B52">2018</xref>) found that out of 12 ratings scales, a vocalic guise manipulation yielded only three significant differences, compared to eight significant differences for both speaker region and speaker gender and eleven significant differences for speaker ethnicity. In other words, while the effect revealed by the intonation model is modest, it is possible that this is a general property of phonetic guise manipulations, as well as an artifact of the interconnected nature of suprasegmental features in general and the resulting challenges in isolating them from one another.</p>
</sec>
<sec>
<title>5.3. Incrementality</title>
<p>These findings support the notion that listeners attend to phonetic detail in constructing social meanings of sociophonetic variation, given that listener ratings of blackness for L+H* increased stepwise as L+H* pitch accents became more phonetically extreme. In other words, there is some evidence that listeners map continuous social meanings to continuous variation, supporting our incrementality hypothesis; contra Podesva&#8217;s (<xref ref-type="bibr" rid="B38">2011</xref>) phonetic salience hypothesis, these findings suggest that greater social meanings are not only attached to phonetic outliers, but also to phonetically intermediate realizations of L+H* pitch accents. This research also sheds light on how phonetic salience works in context. While the intonation analysis found a jump between manipulation steps 2 and 3 in listener ratings of blackness for L+H* phrases, the analysis of the L+H* voice quality model&#8217;s random effects found considerable differences in step 1 ratings across L+H* phrases. That is, for some stimuli smaller differences in intonation were sufficient to trigger higher listener ratings of blackness; for others listener ratings of blackness only increased with larger differences in intonation. Thus, just as context shapes the social meaning of a variant&#8217;s presence or absence (<xref ref-type="bibr" rid="B6">Campbell-Kibler, 2009</xref>; <xref ref-type="bibr" rid="B13">Gumperz, 1982</xref>; <xref ref-type="bibr" rid="B26">Leach et al., 2016</xref>; <xref ref-type="bibr" rid="B33">Pharao, Maegaard, M&#248;ller, &amp; Kristiansen, 2014</xref>), context also shapes the way that phonetic detail affects social meanings.</p>
</sec>
<sec>
<title>5.4. Open-guise versus matched-guise technique</title>
<p>These findings expand our understanding of methods for probing language attitudes, countering the received wisdom in MGT research that these tasks only work if listeners believe they are judging different speakers (<xref ref-type="bibr" rid="B12">Giles &amp; Billings, 2004</xref>). This work expands on the findings of Soukup (<xref ref-type="bibr" rid="B43">2013</xref>) in demonstrating additional support for the OGT: Listeners were aware that they were hearing the same speaker, but the guise manipulation nevertheless yielded a difference in listener responses. Soukup (<xref ref-type="bibr" rid="B43">2013</xref>) finds that the OGT yielded larger effects than the MGT on &#8216;superiority&#8217; scales (<xref ref-type="bibr" rid="B53">Zahn &amp; Hopper, 1985</xref>), while the MGT yielded larger effects on &#8216;social attractiveness&#8217; scales. However, her comparison did not address socio-indexical traits like ethnicity that fall outside the superiority-versus-social-attractiveness rubric, but which nevertheless form an important part of listeners&#8217; awareness of language variation (e.g., <xref ref-type="bibr" rid="B14">Hay &amp; Drager, 2010</xref>; <xref ref-type="bibr" rid="B22">Koops, Gentry, &amp; Pantos, 2008</xref>; <xref ref-type="bibr" rid="B32">Niedzielski, 1999</xref>). Although it is impossible to determine how the results of this study would compare to a hypothetical companion MGT (as the MGT simply wouldn&#8217;t work with such a recognizable stimulus speaker)&#8212;and the small effect size we found suggests that a hypothetical companion MGT could yield larger effects&#8212;the present study indicates that a socio-indexical trait, ethnicity, <italic>can</italic> also work in an OGT context.</p>
<p>Moreover, whereas stimulus speakers in typical MGTs are anonymous to listeners, representing blank attitudinal canvases save for small bits of contextual information provided via stimulus text and/or explicit labels, listeners in this study likely had salient prior impressions of President Obama and his racialized speech. The finding that the guise manipulation affected listener perceptions of Obama&#8217;s blackness is even <italic>more</italic> persuasive against that backdrop. Indeed, among the qualitative questionnaire codes that failed to significantly improve the model was ObamaIsBlack (see Appendix B); that is, we found no evidence that listener perceptions of blackness were affected by whether listeners found it difficult to rate Obama as &#8216;sounding white.&#8217;</p>
<p>Although the OGT worked in the present study, we caution readers against the assumption that the OGT will necessarily apply to any context, feature, or trait. First, while both Soukup&#8217;s study and the present study intentionally violated the assumption that listeners should believe they are judging different speakers, in both studies listeners were not told which <italic>feature</italic> was manipulated; we argue that this remains an important element of methodological opacity in speaker evaluation tasks. It is likely that doing so would produce rather different results than if listeners are not informed, especially for those few sociolinguistic variables that attract public commentary. Indeed, only 20% of the listeners in the current study reported that they could detect the guise manipulation (DetectManip, Appendix B), and this failed to significantly improve the model; this is helped by the fact that, aside from high rising terminal (<xref ref-type="bibr" rid="B50">Tyler, 2015</xref>), intonational variation is generally not a subject of public commentary in American English.</p>
<p>Second, we argue that there remain contexts in which it is important to conceal the fact that the same speaker is behind both or all guises. While the majority of speaker evaluation tasks involve cognitive and/or affective responses, we predict that tasks involving behavioral responses (e.g., making a hiring decision) are likelier to hinge on listeners believing they are hearing different speakers. For example, if the landlords in Purnell et al. (<xref ref-type="bibr" rid="B39">1999</xref>) knew they were hearing John Baugh in multiple guises, they might have been on their &#8216;best behavior&#8217; to avoid prosecution under the Fair Housing Act.</p>
<p>Third, we argue that the use of an OGT rather than MGT approach must be justified by a plausible style-shifting context. For example, this task relied on listeners&#8217; awareness of President Obama&#8217;s style-shifting to sound more black in some contexts and less black in others (<xref ref-type="bibr" rid="B1">Alim &amp; Smitherman, 2012</xref>); as mentioned above, it is conceivable that listeners&#8217; awareness in this respect was facilitated by their generally positive attitude toward Obama&#8217;s presidency making them more likely to hear Obama&#8217;s public speaking. In a similar justification of a plausible style-shifting context, Soukup (<xref ref-type="bibr" rid="B43">2013</xref>) relied on her observation that speakers routinely shift between standard and dialectal Austrian German in stylistic practice. If a speaker evaluation task involves styles that do not coexist in stylistic practice &#8216;in the wild&#8217; (e.g., the same speaker commanding both an L1 and an L2 accent), the OGT is not likely to work.</p>
<p>Caveats about the OGT notwithstanding, it is clear that traditional approaches to linguistic perception do not give listeners enough credit for being aware of style-shifting; indeed, explicit public awareness of style-shifting (e.g., <xref ref-type="bibr" rid="B31">Meraji, 2013</xref>) indicates that listeners may be willing to accept reacting to the same speaker using different features, styles, or languages. Future research should explore the extent to which style-shifting <italic>itself</italic>, not just the individual styles involved in shifting, affects listeners&#8217; judgments of speakers.</p>
</sec>
</sec>
<sec>
<title>6. Conclusion</title>
<p>The current study examined listener ratings of phonetically manipulated speech to test whether listeners were sensitive to such manipulations in the process of making judgments about speaker ethnicity. Regression models indicated that listeners systematically judged a familiar speaker as &#8216;sounding blacker&#8217; when exposed to more extreme F0 manipulations of both the peak and valley of L+H* pitch accents. This effect was mediated by incrementality, with more extreme L+H* pitch accents mapping to greater perceptions of blackness&#8212;albeit with an effect size that suggests caution in generalizing these results. Results of post-hoc testing also reveal that a number of voice quality features appear to also be involved in these judgments. In particular, speech rate, peak delay, HNR, and jitter also appear to influence listener judgments, though the salience of voice quality features may be mediated by the presence versus absence of L+H* pitch accents.</p>
<p>These results have important implications for future work examining both intonational variation from a formal perspective as well as sociophonetic studies on ethnic identification. The finding that listeners seem to attune differently to H* versus L+H* pitch accents in ethnicity judgments and that these perceptions are influenced by phonetic factors provides further motivation for studies that examine intonation from both a phonological and a phonetic perspective. Additionally, the finding that listener perceptions of ethnicity may be manipulated by alterations in F0 provides important context for studies that aim to isolate the phonetic features that may trigger listener judgments of ethnicity. This is especially important given the large body of work on linguistic profiling and discrimination and may provide additional resources for linguists who aim to describe and address racial inequality. Finally, these results indicate that listeners&#8217; sociolinguistic perceptions are sensitive to the magnitude of the input, a finding that indicates promising directions for research in language attitudes and sociolinguistic cognition.</p>
</sec>
<sec sec-type="supplementary-material">
<title>Additional Files</title>
<p>The additional files for this article can be found as follows:</p>
<supplementary-material id="S1" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.229.s1">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-11-229-s1.pdf">labphon-11-229-s1.pdf</inline-supplementary-material>]-->
<label>Appendix A</label>
<caption>
<p>Questionnaire. DOI: <uri>https://doi.org/10.5334/labphon.229.s1</uri></p>
</caption>
</supplementary-material>
<supplementary-material id="S2" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.229.s2">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-11-229-s2.pdf">labphon-11-229-s2.pdf</inline-supplementary-material>]-->
<label>Appendix B</label>
<caption>
<p>Questionnaire qualitative codes. DOI: <uri>https://doi.org/10.5334/labphon.229.s2</uri></p>
</caption>
</supplementary-material>
<supplementary-material id="S3" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.5334/labphon.229.s3">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-11-229-s3.pdf">labphon-11-229-s3.pdf</inline-supplementary-material>]-->
<label>Appendix C</label>
<caption>
<p>Voice quality model summaries. DOI: <uri>https://doi.org/10.5334/labphon.229.s3</uri></p>
</caption>
</supplementary-material>
</sec>
</body>
<back>
<fn-group>
<fn id="n1"><p>Listeners were intentionally not provided guidance on how to interpret this question, because earlier ethnic identification studies allowed for speakers to answer with their own conceptualizations of race and ethnicity (cf. <xref ref-type="bibr" rid="B48">Thomas &amp; Reaser, 2004</xref>). Since one of the aims of the current study was to test for incrementality in ethnic judgments, it was important that listeners&#8217; judgments were shaped by their own state of knowledge about ethnolinguistic patterning of intonational variation.</p></fn>
<fn id="n2"><p>Portions of this data appeared in print in the University of Pennsylvania Working Papers, Selected Papers from NWAV46, as &#8220;How black does Obama sound now?: Testing listener judgments of intonation in incrementally manipulated speech.&#8221;</p></fn>
<fn id="n3"><p>Though some scholars have posited that the intonational phonological inventory of AAL may differ from that of MAE, it is still considered a reliable method for analyzing intonation in AAL, at least until researchers further investigate development of an AAL ToBI system (<xref ref-type="bibr" rid="B15">Holliday, 2016</xref>; <xref ref-type="bibr" rid="B16">Thomas, 2015</xref>).</p></fn>
<fn id="n4"><p>Whereas Podesva&#8217;s phonetic salience hypothesis applies only to phonetic outliers, our incrementality hypothesis applies across the &#8216;axis of phonetic variation&#8217;; the latter can thus be considered a stronger form of the former.</p></fn>
<fn id="n5"><p>An anonymous reviewer expresses doubt that this significant effect &#8220;would reliably reappear as an important factor&#8221; in a replication of the present study; we agree that this is an empirical question.</p></fn>
<fn id="n6"><p>Phrase speech rate, peak delay, and vowel duration are prosodic features, not voice quality features, but for the sake of brevity we refer to the entire set as voice quality features.</p></fn>
<fn id="n7"><p>Thanks to an editor for pointing this out.</p></fn>
</fn-group>
<ack>
<title>Acknowledgements</title>
<p>The authors wish to express their thanks to Paul Reed for comments on the study design. We would also like to thank the audiences at New Ways of Analyzing Variation (NWAV46) and Sociolinguistics Symposium 22, as well as anonymous reviewers for their helpful feedback. Thanks also to our listeners.</p>
</ack>
<sec>
<title>Competing Interests</title>
<p>The authors have no competing interests to declare.</p>
</sec>
<ref-list>
<ref id="B1"><label>1</label><mixed-citation publication-type="book"><string-name><surname>Alim</surname>, <given-names>H. S.</given-names></string-name>, &amp; <string-name><surname>Smitherman</surname>, <given-names>G.</given-names></string-name> (<year>2012</year>). <source>Articulate While Black: Barack Obama, Language, and Race in the U.S</source>. <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</mixed-citation></ref>
<ref id="B2"><label>2</label><mixed-citation publication-type="journal"><string-name><surname>Beckman</surname>, <given-names>M. E.</given-names></string-name>, &amp; <string-name><surname>Ayers-Elam</surname>, <given-names>G.</given-names></string-name> (<year>1997</year>). <article-title>Guidelines for ToBI labelling</article-title>, version 3.0.</mixed-citation></ref>
<ref id="B3"><label>3</label><mixed-citation publication-type="webpage"><string-name><surname>Boersma</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Weenink</surname>, <given-names>D.</given-names></string-name> (<year>2015</year>). <article-title>Praat</article-title> (Version 5.4.01) [phonetic analysis software]. Available from <uri>http://www.fon.hum.uva.nl/praat/</uri></mixed-citation></ref>
<ref id="B4"><label>4</label><mixed-citation publication-type="confproc"><string-name><surname>Burdin</surname>, <given-names>R.</given-names></string-name> (<year>2015</year>). <article-title>Phonological and phonetic variation in list intonation in Jewish English</article-title>. <conf-name>Paper presented at NWAV 44</conf-name>. <conf-loc>Toronto</conf-loc>. DOI: <pub-id pub-id-type="doi">10.21437/SpeechProsody.2014-175</pub-id></mixed-citation></ref>
<ref id="B5"><label>5</label><mixed-citation publication-type="confproc"><string-name><surname>Burdin</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Holliday</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Reed</surname>, <given-names>P.</given-names></string-name> (<year>2018</year>). <article-title>Rising above the standard: Variation in L+H* contour use across 5 varieties of American English</article-title>. <conf-name>Paper presented at Speech Prosody</conf-name>. DOI: <pub-id pub-id-type="doi">10.21437/SpeechProsody.2018-72</pub-id></mixed-citation></ref>
<ref id="B6"><label>6</label><mixed-citation publication-type="journal"><string-name><surname>Campbell-Kibler</surname>, <given-names>K.</given-names></string-name> (<year>2009</year>). <article-title>The nature of sociolinguistic perception</article-title>. <source>Language Variation and Change</source>, <volume>21</volume>(<issue>1</issue>), <fpage>135</fpage>&#8211;<lpage>156</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0954394509000052</pub-id></mixed-citation></ref>
<ref id="B7"><label>7</label><mixed-citation publication-type="book"><string-name><surname>Clopper</surname>, <given-names>C. G.</given-names></string-name> (<year>2010</year>). <chapter-title>Phonetic detail, linguistic experience, and the classification of regional language varieties in the United States</chapter-title>. In <string-name><given-names>D. R.</given-names> <surname>Preston</surname></string-name> &amp; <string-name><given-names>N.</given-names> <surname>Niedzielski</surname></string-name> (Eds.), <source>A reader in sociophonetics</source> (pp. <fpage>203</fpage>&#8211;<lpage>222</lpage>). <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>.</mixed-citation></ref>
<ref id="B8"><label>8</label><mixed-citation publication-type="journal"><string-name><surname>Clopper</surname>, <given-names>C. G.</given-names></string-name>, &amp; <string-name><surname>Pisoni</surname>, <given-names>D. B.</given-names></string-name> (<year>2007</year>). <article-title>Free classification of regional dialects of American English</article-title>. <source>Journal of Phonetics</source>, <volume>35</volume>(<issue>3</issue>), <fpage>421</fpage>&#8211;<lpage>438</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2006.06.001</pub-id></mixed-citation></ref>
<ref id="B9"><label>9</label><mixed-citation publication-type="journal"><string-name><surname>D&#8217;Onofrio</surname>, <given-names>A.</given-names></string-name> (<year>2018</year>). <article-title>Personae and phonetic detail in sociolinguistic signs</article-title>. <source>Language in Society</source>, <volume>47</volume>(<issue>4</issue>), <fpage>513</fpage>&#8211;<lpage>539</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0047404518000581</pub-id></mixed-citation></ref>
<ref id="B10"><label>10</label><mixed-citation publication-type="journal"><string-name><surname>Foulkes</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Docherty</surname>, <given-names>G.</given-names></string-name> (<year>2006</year>). <article-title>The social life of phonetics and phonology</article-title>. <source>Journal of Phonetics</source>, <volume>34</volume>(<issue>4</issue>), <fpage>409</fpage>&#8211;<lpage>438</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2005.08.002</pub-id></mixed-citation></ref>
<ref id="B11"><label>11</label><mixed-citation publication-type="book"><string-name><surname>Foulkes</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Docherty</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Khattab</surname>, <given-names>G.</given-names></string-name>, &amp; <string-name><surname>Yaeger-Dror</surname>, <given-names>M.</given-names></string-name> (<year>2010</year>). <chapter-title>Sound judgments: Perception of indexical features in children&#8217;s speech</chapter-title>. In <string-name><given-names>D. R.</given-names> <surname>Preston</surname></string-name> &amp; <string-name><given-names>N.</given-names> <surname>Niedzielski</surname></string-name> (Eds.), <source>A reader in sociophonetics</source> (pp. <fpage>327</fpage>&#8211;<lpage>356</lpage>). <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>.</mixed-citation></ref>
<ref id="B12"><label>12</label><mixed-citation publication-type="book"><string-name><surname>Giles</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>Billings</surname>, <given-names>A. C.</given-names></string-name> (<year>2004</year>). <chapter-title>Assessing language attitudes: Speaker evaluation studies</chapter-title>. In <string-name><given-names>A.</given-names> <surname>Davies</surname></string-name> &amp; <string-name><given-names>C.</given-names> <surname>Elder</surname></string-name> (Eds.), <source>The handbook of applied linguistics</source> (pp. <fpage>187</fpage>&#8211;<lpage>209</lpage>). <publisher-loc>Malden, MA</publisher-loc>: <publisher-name>Blackwell</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1002/9780470757000.ch7</pub-id></mixed-citation></ref>
<ref id="B13"><label>13</label><mixed-citation publication-type="book"><string-name><surname>Gumperz</surname>, <given-names>J. J.</given-names></string-name> (<year>1982</year>). <source>Discourse strategies</source>. <publisher-loc>Cambridge</publisher-loc>: <publisher-name>Cambridge University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1017/CBO9780511611834</pub-id></mixed-citation></ref>
<ref id="B14"><label>14</label><mixed-citation publication-type="journal"><string-name><surname>Hay</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Drager</surname>, <given-names>K.</given-names></string-name> (<year>2010</year>). <article-title>Stuffed toys and speech perception</article-title>. <source>Linguistics</source>, <volume>48</volume>(<issue>4</issue>), <fpage>865</fpage>&#8211;<lpage>892</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/ling.2010.027</pub-id></mixed-citation></ref>
<ref id="B15"><label>15</label><mixed-citation publication-type="thesis"><string-name><surname>Holliday</surname>, <given-names>N.</given-names></string-name> (<year>2016</year>). <source>Intonational Variation, Linguistic Style, and the Black/Biracial Experience</source>. (Doctoral dissertation), <publisher-name>New York University</publisher-name>.</mixed-citation></ref>
<ref id="B16"><label>16</label><mixed-citation publication-type="confproc"><string-name><surname>Holliday</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Jaggers</surname>, <given-names>Z. S.</given-names></string-name> (<year>2015</year>). <article-title>Influence of suprasegmental features on perceived ethnicity of American politicians</article-title>. <conf-name>Paper presented at 18th International Congress of Phonetic Sciences</conf-name>, <conf-loc>Glasgow, Scotland</conf-loc>.</mixed-citation></ref>
<ref id="B17"><label>17</label><mixed-citation publication-type="journal"><string-name><surname>Jacewicz</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Fox</surname>, <given-names>R. A.</given-names></string-name> (<year>2018</year>). <article-title>The old, the new, and the in&#8208;between: Preadolescents&#8217; use of stylistic variation in speech in projecting their own identity in a culturally changing environment</article-title>. <source>Developmental Science</source>. DOI: <pub-id pub-id-type="doi">10.1111/desc.12722</pub-id></mixed-citation></ref>
<ref id="B18"><label>18</label><mixed-citation publication-type="journal"><string-name><surname>Jun</surname>, <given-names>S.-A.</given-names></string-name>, &amp; <string-name><surname>Foreman</surname>, <given-names>C.</given-names></string-name> (<year>1996</year>). <article-title>Boundary tones and focus realization in African American English intonations</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>100</volume>(<issue>4</issue>), 2826. DOI: <pub-id pub-id-type="doi">10.1121/1.416648</pub-id></mixed-citation></ref>
<ref id="B19"><label>19</label><mixed-citation publication-type="book"><string-name><surname>Kaplan</surname>, <given-names>R.</given-names></string-name> (Writer). (<year>2016</year>). <publisher-loc>Obamas share Super Bowl traditions with Gayle King</publisher-loc>: <publisher-name>CBS News</publisher-name>.</mixed-citation></ref>
<ref id="B20"><label>20</label><mixed-citation publication-type="confproc"><string-name><surname>Keating</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Garellek</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Kreiman</surname>, <given-names>J.</given-names></string-name> (<year>2015</year>). <article-title>Acoustic properties of different kinds of creaky voice</article-title>. <conf-name>Paper presented at 18th International Congress of Phonetic Sciences</conf-name>, <conf-loc>Glasgow, Scotland</conf-loc>. DOI: <pub-id pub-id-type="doi">10.1017/S0025100315000286</pub-id></mixed-citation></ref>
<ref id="B21"><label>21</label><mixed-citation publication-type="book"><string-name><surname>Kendall</surname>, <given-names>T.</given-names></string-name> (<year>2013</year>). <source>Speech rate, pause and sociolinguistic variation: Studies in corpus sociophonetics</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Palgrave Macmillan</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1057/9781137291448</pub-id></mixed-citation></ref>
<ref id="B22"><label>22</label><mixed-citation publication-type="journal"><string-name><surname>Koops</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Gentry</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Pantos</surname>, <given-names>A.</given-names></string-name> (<year>2008</year>). <article-title>The effect of perceived speaker age on the perception of PIN and PEN vowels in Houston, Texas</article-title>. <source>University of Pennsylvania Working Papers in Linguistics</source>, <volume>14</volume>(<issue>2</issue>).</mixed-citation></ref>
<ref id="B23"><label>23</label><mixed-citation publication-type="webpage"><string-name><surname>Kuznetsova</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Brockhoff</surname>, <given-names>B.</given-names></string-name>, &amp; <string-name><surname>Christensen</surname>, <given-names>H. B.</given-names></string-name> (<year>2016</year>). <article-title>lmerTest</article-title> (Version 2.0-33) [R package]. Available from <uri>https://CRAN.R-project.org/package=lmerTest</uri></mixed-citation></ref>
<ref id="B24"><label>24</label><mixed-citation publication-type="book"><string-name><surname>Labov</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Ash</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Boberg</surname>, <given-names>C.</given-names></string-name> (<year>2006</year>). <source>The atlas of North American English: Phonetics, phonology and sound change</source>. <publisher-loc>Berlin</publisher-loc>: <publisher-name>Mouton de Gruyter</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1515/9783110167467</pub-id></mixed-citation></ref>
<ref id="B25"><label>25</label><mixed-citation publication-type="journal"><string-name><surname>Labov</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Ash</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Ravindranath</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Weldon</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Baranowski</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Nagy</surname>, <given-names>N.</given-names></string-name> (<year>2011</year>). <article-title>Properties of the sociolinguistic monitor</article-title>. <source>Journal of Sociolinguistics</source>, <volume>15</volume>(<issue>4</issue>), <fpage>431</fpage>&#8211;<lpage>463</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/j.1467-9841.2011.00504.x</pub-id></mixed-citation></ref>
<ref id="B26"><label>26</label><mixed-citation publication-type="journal"><string-name><surname>Leach</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Watson</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Gnevsheva</surname>, <given-names>K.</given-names></string-name> (<year>2016</year>). <article-title>Perceptual dialectology in northern England: Accent recognition, geographical proximity and cultural prominence</article-title>. <source>Journal of Sociolinguistics</source>, <volume>20</volume>(<issue>2</issue>), <fpage>192</fpage>&#8211;<lpage>211</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/josl.12178</pub-id></mixed-citation></ref>
<ref id="B27"><label>27</label><mixed-citation publication-type="journal"><string-name><surname>Lefcheck</surname>, <given-names>J. S.</given-names></string-name> (<year>2016</year>). <article-title>piecewiseSEM: Piecewise structural equation modelling in R for ecology, evolution, and systematics</article-title>. <source>Methods in Ecology and Evolution</source>, <volume>7</volume>(<issue>5</issue>), <fpage>573</fpage>&#8211;<lpage>579</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/2041-210X.12512</pub-id></mixed-citation></ref>
<ref id="B28"><label>28</label><mixed-citation publication-type="book"><string-name><surname>Loman</surname>, <given-names>B.</given-names></string-name> (<year>1975</year>). <chapter-title>Prosodic patterns in a Negro American dialect</chapter-title>. In <string-name><given-names>H.</given-names> <surname>Ringbom</surname></string-name>, <string-name><given-names>A.</given-names> <surname>Ingberg</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Norrman</surname></string-name>, <string-name><given-names>K.</given-names> <surname>Nyholm</surname></string-name>, <string-name><given-names>R.</given-names> <surname>Westman</surname></string-name> &amp; <string-name><given-names>K.</given-names> <surname>Wikberg</surname></string-name> (Eds.), <source>Style and text: Studies presented to Nils Erik Enkvist</source> (pp. <fpage>219</fpage>&#8211;<lpage>242</lpage>). <publisher-loc>Stockholm</publisher-loc>: <publisher-name>Spr&#229;kf&#246;rlaget Skriptor AB</publisher-name>.</mixed-citation></ref>
<ref id="B29"><label>29</label><mixed-citation publication-type="journal"><string-name><surname>McLarty</surname>, <given-names>J.</given-names></string-name> (<year>2018</year>). <article-title>African American Language and European American English intonation variation over time in the American South</article-title>. <source>American Speech</source>, <volume>93</volume>(<issue>1</issue>), <fpage>32</fpage>&#8211;<lpage>78</lpage>. DOI: <pub-id pub-id-type="doi">10.1215/00031283-6904032</pub-id></mixed-citation></ref>
<ref id="B30"><label>30</label><mixed-citation publication-type="confproc"><string-name><surname>McLarty</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Vaughn</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Kendall</surname>, <given-names>T.</given-names></string-name> (<year>2017</year>). <article-title>Acoustic correlates of perceived prosodic prominence in African American English and European American English</article-title>. <conf-name>Paper presented at NWAV 46</conf-name>, <conf-loc>Madison, Wisconsin</conf-loc>.</mixed-citation></ref>
<ref id="B31"><label>31</label><mixed-citation publication-type="webpage"><string-name><surname>Meraji</surname>, <given-names>S. M.</given-names></string-name> (<year>2013</year>). <article-title>Why Chaucer said &#8216;ax&#8217; instead of &#8216;ask,&#8217; and why some still do</article-title>. <source>National Public Radio</source>. Retrieved from <uri>https://www.npr.org/sections/codeswitch/2013/12/03/248515217/why-chaucer-said-ax-instead-of-ask-and-why-some-still-do</uri></mixed-citation></ref>
<ref id="B32"><label>32</label><mixed-citation publication-type="journal"><string-name><surname>Niedzielski</surname>, <given-names>N.</given-names></string-name> (<year>1999</year>). <article-title>The effect of social information on the perception of sociolinguistic variables</article-title>. <source>Journal of Language and Social Psychology</source>, <volume>18</volume>(<issue>1</issue>), <fpage>62</fpage>&#8211;<lpage>85</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0261927X99018001005</pub-id></mixed-citation></ref>
<ref id="B33"><label>33</label><mixed-citation publication-type="journal"><string-name><surname>Pharao</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Maegaard</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>M&#248;ller</surname>, <given-names>J. S.</given-names></string-name>, &amp; <string-name><surname>Kristiansen</surname>, <given-names>T.</given-names></string-name> (<year>2014</year>). <article-title>Indexical meanings of [s+] among Copenhagen youth: Social perception of a phonetic variant in different prosodic contexts</article-title>. <source>Language in Society</source>, <volume>43</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>31</lpage>. DOI: <pub-id pub-id-type="doi">10.1017/S0047404513000857</pub-id></mixed-citation></ref>
<ref id="B34"><label>34</label><mixed-citation publication-type="thesis"><string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name> (<year>1980</year>). <source>The phonology and phonetics of English intonation</source>. (Doctoral dissertation), <publisher-name>Massachusetts Institute of Technology</publisher-name>, <publisher-loc>Cambridge, MA</publisher-loc>.</mixed-citation></ref>
<ref id="B35"><label>35</label><mixed-citation publication-type="journal"><string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name> (<year>2016</year>). <article-title>Phonological representation: Beyond abstract versus episodic</article-title>. <source>Annual Review of Linguistics</source>, <volume>2</volume>(<issue>1</issue>). DOI: <pub-id pub-id-type="doi">10.1146/annurev-linguistics-030514-125050</pub-id></mixed-citation></ref>
<ref id="B36"><label>36</label><mixed-citation publication-type="book"><string-name><surname>Pierrehumbert</surname>, <given-names>J. B.</given-names></string-name>, &amp; <string-name><surname>Hirschberg</surname>, <given-names>J.</given-names></string-name> (<year>1990</year>). <chapter-title>The meaning of intonational contours in the interpretation of discourse</chapter-title>. In <string-name><given-names>P. R.</given-names> <surname>Cohen</surname></string-name>, <string-name><given-names>J.</given-names> <surname>Morgan</surname></string-name> &amp; <string-name><given-names>M.</given-names> <surname>Pollack</surname></string-name> (Eds.), <source>Intentions in communication</source> (pp. <fpage>271</fpage>&#8211;<lpage>311</lpage>). <publisher-loc>Cambridge, MA</publisher-loc>: <publisher-name>MIT Press</publisher-name>.</mixed-citation></ref>
<ref id="B37"><label>37</label><mixed-citation publication-type="journal"><string-name><surname>Plichta</surname>, <given-names>B.</given-names></string-name>, &amp; <string-name><surname>Preston</surname>, <given-names>D. R.</given-names></string-name> (<year>2005</year>). <article-title>The /ay/s have it: The perception of /ay/ as a North-South stereotype in US English</article-title>. <source>Acta Linguistica Hafniensia</source>, <volume>37</volume>, <fpage>243</fpage>&#8211;<lpage>285</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/03740463.2005.10416086</pub-id></mixed-citation></ref>
<ref id="B38"><label>38</label><mixed-citation publication-type="journal"><string-name><surname>Podesva</surname>, <given-names>R. J.</given-names></string-name> (<year>2011</year>). <article-title>Salience and the social meaning of declarative contours</article-title>. <source>Journal of English Linguistics</source>, <volume>39</volume>(<issue>3</issue>), <fpage>233</fpage>&#8211;<lpage>264</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0075424211405161</pub-id></mixed-citation></ref>
<ref id="B39"><label>39</label><mixed-citation publication-type="journal"><string-name><surname>Purnell</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Idsardi</surname>, <given-names>W. J.</given-names></string-name>, &amp; <string-name><surname>Baugh</surname>, <given-names>J.</given-names></string-name> (<year>1999</year>). <article-title>Perceptual and phonetic experiments on American English dialect identification</article-title>. <source>Journal of Language and Social Psychology</source>, <volume>18</volume>(<issue>1</issue>), <fpage>10</fpage>&#8211;<lpage>30</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0261927X99018001002</pub-id></mixed-citation></ref>
<ref id="B40"><label>40</label><mixed-citation publication-type="webpage"><collab>R Core Team</collab>. (<year>2018</year>). <chapter-title>R: A language and environment for statistical computing</chapter-title> (Version 3.5.2). <publisher-loc>Vienna</publisher-loc>. Available from <uri>https://www.R-project.org/</uri></mixed-citation></ref>
<ref id="B41"><label>41</label><mixed-citation publication-type="thesis"><string-name><surname>Reed</surname>, <given-names>P.</given-names></string-name> (<year>2016</year>). <source>Sounding Appalachian: /aI/ Monophthongization, Rising Pitch Accents, and Rootedness</source>. (Doctoral dissertation), <publisher-name>University of South Carolina</publisher-name>.</mixed-citation></ref>
<ref id="B42"><label>42</label><mixed-citation publication-type="journal"><string-name><surname>Satterthwaite</surname>, <given-names>F. E.</given-names></string-name> (<year>1946</year>). <article-title>An Approximate Distribution of Estimates of Variance Components</article-title>. <source>Biometrics Bulletin</source>, <volume>2</volume>(<issue>6</issue>), <fpage>110</fpage>&#8211;<lpage>114</lpage>. DOI: <pub-id pub-id-type="doi">10.2307/3002019</pub-id></mixed-citation></ref>
<ref id="B43"><label>43</label><mixed-citation publication-type="confproc"><string-name><surname>Soukup</surname>, <given-names>B.</given-names></string-name> (<year>2013</year>). <article-title>&#8216;Matched guise technique&#8217; vs. &#8216;Open guise technique&#8217; in the elicitation of language attitudes: Insights from a comparative study</article-title>. <conf-name>Paper presented at ExAPP 2</conf-name>, <conf-loc>Copenhagen</conf-loc>.</mixed-citation></ref>
<ref id="B44"><label>44</label><mixed-citation publication-type="journal"><string-name><surname>Szakay</surname>, <given-names>A.</given-names></string-name> (<year>2012</year>). <article-title>Voice quality as a marker of ethnicity in New Zealand: From acoustics to perception</article-title>. <source>Journal of Sociolinguistics</source>, <volume>16</volume>(<issue>3</issue>), <fpage>382</fpage>&#8211;<lpage>397</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/j.1467-9841.2012.00537.x</pub-id></mixed-citation></ref>
<ref id="B45"><label>45</label><mixed-citation publication-type="journal"><string-name><surname>Tarone</surname>, <given-names>E.</given-names></string-name> (<year>1973</year>). <article-title>Aspects of intonation in Black English</article-title>. <source>American Speech</source>, <volume>48</volume>(<issue>1/2</issue>), <fpage>29</fpage>&#8211;<lpage>36</lpage>. DOI: <pub-id pub-id-type="doi">10.2307/3087890</pub-id></mixed-citation></ref>
<ref id="B46"><label>46</label><mixed-citation publication-type="book"><string-name><surname>Thomas</surname>, <given-names>E. R.</given-names></string-name> (<year>2011</year>). <source>Sociophonetics: An introduction</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Palgrave Macmillan</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1007/978-1-137-28561-4</pub-id></mixed-citation></ref>
<ref id="B47"><label>47</label><mixed-citation publication-type="book"><string-name><surname>Thomas</surname>, <given-names>E. R.</given-names></string-name> (<year>2015</year>). <chapter-title>Prosodic features of African American English</chapter-title>. In <string-name><given-names>S.</given-names> <surname>Lanehart</surname></string-name> (Ed.), <source>The Oxford handbook of African American Language</source> (pp. <fpage>420</fpage>&#8211;<lpage>438</lpage>). <publisher-loc>Oxford</publisher-loc>: <publisher-name>Oxford University Press</publisher-name>.</mixed-citation></ref>
<ref id="B48"><label>48</label><mixed-citation publication-type="journal"><string-name><surname>Thomas</surname>, <given-names>E. R.</given-names></string-name>, &amp; <string-name><surname>Reaser</surname>, <given-names>J.</given-names></string-name> (<year>2004</year>). <article-title>Delimiting perceptual cues used for the ethnic labeling of African American and European American voices</article-title>. <source>Journal of Sociolinguistics</source>, <volume>8</volume>(<issue>1</issue>), <fpage>54</fpage>&#8211;<lpage>87</lpage>. DOI: <pub-id pub-id-type="doi">10.1111/j.1467-9841.2004.00251.x</pub-id></mixed-citation></ref>
<ref id="B49"><label>49</label><mixed-citation publication-type="confproc"><string-name><surname>Todd</surname>, <given-names>R.</given-names></string-name> (<year>2002</year>). <article-title>Speaker-ethnicity: Attributions based on the use of prosodic cues</article-title>. <conf-name>Paper presented at Speech Prosody</conf-name>, <conf-loc>Aix-en-Provence, France</conf-loc>.</mixed-citation></ref>
<ref id="B50"><label>50</label><mixed-citation publication-type="journal"><string-name><surname>Tyler</surname>, <given-names>J. C.</given-names></string-name> (<year>2015</year>). <article-title>Expanding and mapping the indexical field: Rising pitch, the uptalk stereotype, and perceptual variation</article-title>. <source>Journal of English Linguistics</source>, <volume>43</volume>(<issue>4</issue>), <fpage>284</fpage>&#8211;<lpage>310</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0075424215607061</pub-id></mixed-citation></ref>
<ref id="B51"><label>51</label><mixed-citation publication-type="book"><string-name><surname>Villarreal</surname>, <given-names>D.</given-names></string-name> (<year>2016</year>). <chapter-title>&#8220;Do I sound like a Valley Girl to you?&#8221; Perceptual dialectology and language attitudes in California</chapter-title>. In <string-name><given-names>V.</given-names> <surname>Fridland</surname></string-name>, <string-name><given-names>T.</given-names> <surname>Kendall</surname></string-name>, <string-name><given-names>B.</given-names> <surname>Evans</surname></string-name> &amp; <string-name><given-names>A.</given-names> <surname>Wassink</surname></string-name> (Eds.), <source>Speech in the Western states</source> (vol. <volume>1</volume>, pp. <fpage>55</fpage>&#8211;<lpage>75</lpage>). <publisher-loc>Durham, NC</publisher-loc>: <publisher-name>Duke University Press</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1215/00031283-3772901</pub-id></mixed-citation></ref>
<ref id="B52"><label>52</label><mixed-citation publication-type="journal"><string-name><surname>Villarreal</surname>, <given-names>D.</given-names></string-name> (<year>2018</year>). <article-title>The construction of social meaning: A matched-guise investigation of the California Vowel Shift</article-title>. <source>Journal of English Linguistics</source>, <volume>46</volume>(<issue>1</issue>), <fpage>52</fpage>&#8211;<lpage>78</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0075424217753520</pub-id></mixed-citation></ref>
<ref id="B53"><label>53</label><mixed-citation publication-type="journal"><string-name><surname>Zahn</surname>, <given-names>C. J.</given-names></string-name>, &amp; <string-name><surname>Hopper</surname>, <given-names>R.</given-names></string-name> (<year>1985</year>). <article-title>Measuring language attitudes: The speech evaluation instrument</article-title>. <source>Journal of Language and Social Psychology</source>, <volume>4</volume>(<issue>2</issue>), <fpage>113</fpage>&#8211;<lpage>123</lpage>. DOI: <pub-id pub-id-type="doi">10.1177/0261927X8500400203</pub-id></mixed-citation></ref>
</ref-list>
</back>
</article>