<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.0 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.0/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.0" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology: Journal of the Association for Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/labphon.14</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Analytical Decisions in Intonation Research and the Role of Representations: Lessons from Romani</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>Amalia</given-names>
</name>
<email>a.arvaniti@kent.ac.uk</email>
<xref ref-type="aff" rid="aff-1"/>
</contrib>
</contrib-group>
<aff id="aff-1">English Language and Linguistics, SECL, University of Kent, Cornwallis NW, Canterbury CT2 7NF, UK</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2016-06-30">
<day>30</day>
<month>06</month>
<year>2016</year>
</pub-date>
<volume>7</volume>
<issue>1</issue>
<elocation-id>6</elocation-id>
<supplementary-material id="S1" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.smo"/>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2016 The Author(s)</copyright-statement>
<copyright-year>2016</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.5334/labphon.14/"/>
<abstract>
<p>This paper presents an analysis of the intonational system of Greek Thrace Romani. The analysis serves to highlight the difficulties that spontaneous fieldwork data pose for traditional methods of intonational research largely developed for use with controlled speech elicited in the laboratory or under laboratory-like conditions from educated speakers of standardized languages. It leads to proposing a set of principles and procedures which can help deal with the variability inherent in spontaneous data; these principles and procedures apply particularly to data from less homogeneous speech communities but are relevant for the intonation analysis of any linguistic system. This approach relies on the understanding that autosegmental-metrical representations of intonation are phonological representations, not means of faithfully depicting pitch contours per se. It follows that representations should capture what is contrastive in the intonational system under analysis. In turn, this entails that new categories are posited, taking the meaning of tonal events into account and after due consideration of all legitimate sources of phonetic variation. It is argued that following this procedure allows for more robust analyses and is particularly advantageous when data are highly variable. This view is discussed in light of the analysis of Greek Thrace Romani, and in combination with recent proposals for greater uniformity and phonetic transparency in intonational representations, traits which are said to lead to greater insights in typological and cross-varietal research. It is shown that these goals are not better served by a level of broad phonetic transcription which encodes an arbitrary selection of phonetic variants.</p>
</abstract>
</article-meta>
</front>
<body>
<sec>
<title>1 Introduction</title>
<p>Much of the research in intonation has been laboratory-based. Paradigms for data collection and techniques that are widely accepted in intonational research were originally developed for use with controlled speech elicited from educated speakers of standardized languages. These paradigms and analytical techniques, concisely described in Jun and Fletcher (<xref ref-type="bibr" rid="B68">2014</xref>), have been successfully adopted in the documentation and analysis of a variety of intonational systems that go well beyond the languages for which they were originally developed (see <xref ref-type="bibr" rid="B52">Gussenhoven, 2004</xref>; <xref ref-type="bibr" rid="B65">Jun 2005a</xref>, <xref ref-type="bibr" rid="B67">2014</xref>, for a number of languages analyzed along these lines). These methods have been very useful in determining a number of properties of the systems studied, including prosodic type, levels of phrasing, and tonal inventory.</p>
<p>The data collection paradigms in particular are well suited for the study of mainstream standardized languages as they largely involve the reading aloud of multiple repetitions of specially prepared sentences, short dialogues, or passages. Care is typically taken to use words mostly composed of sonorants in order to minimize microprosodic perturbations, a practice that results in largely smooth pitch tracks (cf. <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>, for advice on this point). Cooperative games and tasks are also used, like the map task (<xref ref-type="bibr" rid="B3">Anderson et al., 1991</xref>), various forms of the discourse completion task (DCT; see, e.g., <xref ref-type="bibr" rid="B29">Borr&#224;s-Comes et al., 2014</xref>, and references therein), and specially designed games (e.g., <xref ref-type="bibr" rid="B110">Swerts et al., 2002</xref>). An example of the widespread use of these data collection paradigms is the Interactive Atlas of Spanish Intonation which includes data from 10 varieties of Spanish elicited using most of the tasks mentioned above (<xref ref-type="bibr" rid="B101">Prieto &amp; Roseano, 2010</xref>). Using such tasks allows for the collection of semi-spontaneous data that still contain a number of controlled parameters. For instance, the words used are comprised mostly of sonorants and may be controlled for other variables like the position of stress or the type of structure elicited (e.g., the original HCRC map task maps contain single nouns, compounds, and noun phrases). In short, these practices result in largely smooth pitch tracks that are mostly uniform both within and across study participants and allow researchers to test specific hypotheses about the role of metrical structure, information structure, or any other parameter that is of interest.</p>
<p>The uniformity of data collected in the laboratory is so extensive that it is often considered a natural feature of human speech (see, e.g., <xref ref-type="bibr" rid="B75">Ladd, 1999</xref>) or at least highly desirable (<xref ref-type="bibr" rid="B117">Xu, 2010</xref>). For this reason it is worth enumerating the facets of similarity researchers have come to expect from intonation data based on characteristics of speech elicited in the laboratory or under laboratory-like conditions, particularly from educated speakers of standardized languages. First, such speakers can read aloud fluently and can do so for multiple repetitions while maintaining a consistent style that is similar across speakers and familiar to all from school.<xref ref-type="fn" rid="n1">1</xref> This in turn means that balanced experimental designs with data that are comparable across speakers are the norm. Even semi-spontaneous tasks such as the map task or the DCT are based on skills that participants are likely to be familiar with, such as map reading and role-playing. Thus even in these less controlled tasks, participants are expected to maintain a consistent speaking style, speech rate, and volume and to follow turn-taking (<xref ref-type="bibr" rid="B104">Sachs et al., 1974</xref>). Further, speakers in the laboratory are likely to be young, educated, and middle class, characteristics that facilitate research in practical ways well: for example, such participants are likely to have healthy voices and use modal phonation (unless a different phonation mode is sociolinguistically appropriate for the community, such as creaky voice in California; <xref ref-type="bibr" rid="B96">Podesva, 2007</xref>; <xref ref-type="bibr" rid="B118">Yuasa, 2010</xref>). The importance of these elements cannot be underestimated, but it becomes apparent only when these conditions are not met (see, e.g., <xref ref-type="bibr" rid="B59">Henrich et al., 2010</xref>, on the expectations arising from research based on samples from Western, Educated, Industrialized, Rich, and Democratic [WEIRD] societies).</p>
<p>As a result of the above, many researchers have come to expect quite uniform data when it comes to intonation, and this expectation is by and large fulfilled, as even in studies that involve varied samples, intonational norms are shared among participants. As an illustration, Ritchart and Arvaniti (<xref ref-type="bibr" rid="B102">2014</xref>) investigated uptalk in California using the map task and eliciting data from a large number of speakers who varied in terms of socioeconomic class, gender, ethnicity, linguistic background, and geographical origin. They found mostly gender-related differences in the frequency and discourse function of uptalk, but few differences relating to form. Chung and Arvaniti (<xref ref-type="bibr" rid="B34">2013</xref>) report data from 15 Seoul Korean speakers, all of whom conformed fully to the intonational patterns of Korean as described in Jun (<xref ref-type="bibr" rid="B66">2005b</xref>), even when performing cycling, a relatively artificial task (<xref ref-type="bibr" rid="B39">Cummins &amp; Port, 1998</xref>).</p>
<p>The limited variability found in such data has led not only to expectations of uniformity in the realization of intonation but has also shaped the field&#8217;s views about the perceived importance of such uniformity. In part the problem relates to the focus on form in much intonational research. This has been so both because attempts at codifying intonational meaning proved too complicated (as in the British School; e.g., <xref ref-type="bibr" rid="B56">Halliday, 1967</xref>; <xref ref-type="bibr" rid="B90">O&#8217;Connor &amp; Arnold, 1973</xref>), and because the consideration of meaning has been limited to basic distinctions such as question vs. statement (for a discussion, see <xref ref-type="bibr" rid="B7">Arvaniti, 2011</xref>; <xref ref-type="bibr" rid="B27">Beckman &amp; Venditti, 2011</xref>). The focus on form coupled with the ability to easily extract pitch tracks and treat them as faithful depictions of intonation appears to have strengthened the view that differences in form, particularly in the alignment of tones with respect to segmental landmarks, are sufficient to establish distinct tonal categories (see <xref ref-type="bibr" rid="B73">Kochanski, 2010</xref>, and <xref ref-type="bibr" rid="B27">Beckman &amp; Venditti, 2011</xref>, for discussion of these practices from different perspectives). As a result of this view, small differences in alignment (and, to a much lesser extent, scaling) can be considered crucial in determining a tonal inventory and are incorporated into phonological analyses (see, e.g., <xref ref-type="bibr" rid="B100">Prieto et al., 2005</xref>, and relevant discussion in <xref ref-type="bibr" rid="B14">Arvaniti et al., 2006a</xref>). In turn, the focus on such differences has led to proposals for a level of intonational representation akin to that of a broad phonetic transcription (<xref ref-type="bibr" rid="B62">Hualde &amp; Prieto, 2016</xref>).</p>
<p>While phonetic detail has taken such an important role in intonation research, concerns have also been voiced that similar intonational phenomena are not analyzed in the same way across languages. This is discussed at some length in Ladd (<xref ref-type="bibr" rid="B76">2008a, pp. 107&#8211;119</xref>), who argues that &#8220;if transcriptions are language-specific, we are left with no theoretically meaningful way to pursue cross-language comparison&#8221; (p. 115). This argument could be interpreted as a plea for more abstract phonological presentations, since abstractions are more likely to converge cross-linguistically, thereby facilitating comparisons. Ladd, however, seems to take the opposite stance: phonetic detail should be faithfully and uniformly represented in intonation research; doing so leads to phonetic transparency which in turn means that crosslinguistic similarities will not be obscured by language-specific representations.</p>
<p>I contend here that Ladd&#8217;s arguments, which focus on the importance of intonational form, implicitly question the legitimacy of abstract phonological representations for intonation and by extension the very legitimacy of intonation as a fully-fledged part of phonological structure (doubts about a fully-fledged phonology of intonation have a long pedigree; see <xref ref-type="bibr" rid="B38">Crystal, 1969</xref>). The fact that intonation is treated differently from other aspects of phonological structure becomes evident if one compares Ladd&#8217;s arguments on intonation with his arguments about segmentals. With respect to segmentals, Ladd (<xref ref-type="bibr" rid="B78">2011</xref>) argues persuasively against the use of a systematic phonetic level and in favour of abstract representations, on the one hand, and of measurable phonetic detail on the other (for similar arguments, see also <xref ref-type="bibr" rid="B94">Pierrehumbert et al., 2000</xref>). However, the systematic phonetic level he argues against when it comes to segmentals is precisely the type of level that Hualde &amp; Prieto (<xref ref-type="bibr" rid="B62">2016</xref>) argue in favour of with respect to intonation; they do so using Ladd&#8217;s own arguments about intonational representations (<xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 107&#8211;130</xref>, <xref ref-type="bibr" rid="B77">2008b</xref>).</p>
<p>The focus on phonetic detail has led to neglecting intonational meaning to an extent that is rather striking when juxtaposed to standard practice in segmental analysis. One cannot imagine a fieldworker deciding that a particular vowel is phonemic in language <italic>x</italic> simply because it sounds similar to a vowel that is phonemic in language <italic>y</italic>. Surely, our hypothetical fieldworker would first wish to consult with native speakers of language <italic>x</italic>, use standard tests such as the presence of (near) minimal pairs, study the role of context in observed variation and <italic>consider the entire phonological system</italic> before establishing the status of that vowel. In other words, she would rely on meaning differences, and context and system-internal observations to reach a decision, not on the precise value of the vowel&#8217;s formants or their similarity to values used in another language.</p>
<p>Of the above criteria, meaning in particular has not featured prominently in descriptions of intonation (but see <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>, for good advice on this point). The importance of meaning becomes evident when data that do not conform to the uniformity assumptions discussed earlier are examined: when faced with variable data, it is difficult if not impossible to rely on similarity of form during analysis. Here, a corpus of Greek Thrace Romani is used to illustrate how analytical decisions can be made in the face of such variable data. Section 2 presents in more detail the reasons for the extensive variability in this corpus; section 3 presents the corpus and the principles used for analysis; section 4 illustrates the use of these principles with respect to stress, tonal inventory, and phrasing in Romani; finally, section 5 discusses the analysis in light of recent calls for surface phonetic representations of intonation (<xref ref-type="bibr" rid="B62">Hualde &amp; Prieto, 2016</xref>), more typological research (<xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 107&#8211;130</xref>, <xref ref-type="bibr" rid="B77">2008b</xref>), and the assumed superiority of laboratory data (<xref ref-type="bibr" rid="B117">Xu, 2010</xref>).</p>
</sec>
<sec>
<title>2 Sources of variability in Greek Thrace Romani</title>
<p>Greek Thrace Romani (henceforth <italic>Romani</italic>) is a Vlax variety of Romani spoken by Muslim Roma in Greek Thrace (<xref ref-type="bibr" rid="B1">Adamou, 2010</xref>; <xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>; <xref ref-type="bibr" rid="B8">Arvaniti &amp; Adamou, 2011</xref>). The Roma are a non-sedentary people who arrived in Europe from North East India approximately 600 years ago. As a people they have long suffered persecution; e.g. an estimated 220,000 Roma died at the hands of the Nazis and their collaborators during WWII (among many, <xref ref-type="bibr" rid="B83">Martins-Heub, 1989</xref>; <xref ref-type="bibr" rid="B112">Tyalglyy, 2009</xref>, and references therein). Possibly as a result of hostile attitudes toward them, the Roma form relatively closed communities that do not easily admit strangers, especially non-Roma.</p>
<p>The above apply to the Greek Roma communities as well. The exact number of Roma in Greece is not certain, as the Greek census of 2011 did not include questions about ethnicity. Estimates range from a minimum of 180,000 to a maximum of 350,000 Roma (or approximately 2.5% of Greece&#8217;s population). The community is not homogeneous: some Greek Roma remain non-sedentary while others are settled in well-known neighbourhoods (e.g., Aghia Varvara in the outskirts of Athens). Communities also differ in terms of religion, with some groups being Muslim and others Christian. In addition, Greek Roma speak a number of Romani dialects; of these, Balkan Romani and Vlax Romani, originating in the Black Sea and Transylvania respectively, are the main ones (<xref ref-type="bibr" rid="B85">Matras, 2002</xref>).</p>
<p>The Romani variety in focus here has been referred to in previous work as <italic>Greek Thrace Xoraxane Romane</italic> (i.e., Turkish Romani) and is a mixture of Turkish and Romani as the name implies (<xref ref-type="bibr" rid="B1">Adamou, 2010</xref>; <xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>). It is recognized as such by the speakers themselves who consider it to be distinct from Romani proper (<xref ref-type="bibr" rid="B1">Adamou, 2010</xref>; <xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>). The data were collected from two communities, Anahoma, close to Komotini, and Drosero, close to Xanthi, both towns in Greek Thrace (see Figure <xref ref-type="fig" rid="F1">1</xref>); each community counts approximately 300 members. Although the distance between Xanthi and Komotini is only 55 km, there are dialectal differences between Anahoma and Drosero: in Anahoma, the Romani variety has mostly Vlax features, while Vlax and Balkan Romani are more mixed in Drosero (<xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>). The differences are partly due to patterns of intermarriage in the two communities (the community in Drosero having closer ties with Roma in Bulgaria than the community in Anahoma). The ambient languages, Turkish and Greek, also exert an influence that adds to variability as is shown in more detail below (<xref ref-type="bibr" rid="B1">Adamou, 2010</xref>).</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Map of Greece, showing Xanthi and Komotini, the Greek Thrace towns in the outskirts of which the Roma communities discussed here are based. Source: By Lencer [CC BY-SA 3.0 (<ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://creativecommons.org/licenses/by-sa/3.0">http://creativecommons.org/licenses/by-sa/3.0</ext-link>)], via Wikimedia Commons.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74464/"/>
</fig>
<p>The speakers of the communities discussed here are trilingual in Romani, Turkish, and Greek. They tend to use Romani at home and within their community; they use Turkish and Greek for trade and other business, transactions with authorities, etc. However, Turkish is rapidly replacing both Romani proper and Turkish Romani (<xref ref-type="bibr" rid="B1">Adamou, 2010</xref>). This is in part due to proximity and close business ties with Turkey, but the trend is also strengthened by the fact that the speakers are Muslim and thus officially considered part of the Greek Muslim minority (which is strongly associated with Turkey). The classification of the Roma as part of the Muslim minority is based on the 1923 Treaty of Lausanne which ended WWI between Turkey and neighbouring states including Greece. The delineation of minorities in the treaty was based on religious rather than ethnic divisions following the practice of the Ottoman Empire.</p>
<p>A corollary of the above is that the Muslim Roma of Greece can be educated either in minority schools, which are Turkish-medium, or in mainstream Greek-medium schools. Although education is compulsory in Greece up to age 15, most of the Roma have at best elementary education. This applies to the speakers in the communities under discussion as well, most of whom have little or no schooling.<xref ref-type="fn" rid="n2">2</xref> Further, because of the minority arrangements, there is no provision in Greek schools for Roma children to learn to read and write in Romani; thus the standard Romani variety used for transnational communication and education purposes in some European countries (<xref ref-type="bibr" rid="B84">Matras, 1999</xref>, <xref ref-type="bibr" rid="B86">2005</xref>) is not known among the Roma in the communities under discussion. As a result of this situation, the reading of controlled sentences in Romani is out of the question, while the translation of sentences from Greek or Turkish is fraught with difficulties as speakers freely mix their three languages. Semi-controlled tasks, though possible as the present corpus demonstrates (see Section 3), must be chosen with care: unschooled speakers are not always comfortable with tasks such as map reading or the role-playing required by DCT, and can be weary of describing images depicting non-naturalistic situations (as happens, for instance, in the Questionnaire for Information Structure or QUIS; <xref ref-type="bibr" rid="B107">Skopeteas et al., 2006</xref>).</p>
<p>The linguistic situation of the communities discussed here has additional consequences for data-gathering. Frequent code-switching and mixing using three languages means that semi-controlled data exhibit more variation than that found in monolingual communities. For example, in elicitation with materials from QUIS, the same speaker would use the Greek word [&#712;kokini] for &#8216;red.F&#8217; and the Romani word [lo&#712;li] in response to prompts immediately following one another. In another QUIS game, one participant would consistently stress the word for &#8216;gorilla&#8217; on the penult, as in Greek, producing [&#611;o&#712;rila], while her interlocutor would vacillate between penultimate and default final stress producing both [&#611;o&#712;rila] and [&#611;ori&#712;la]. Such alternations mean that attempts to control prompts so as to avoid obstruents or elicit specific stress patterns can be easily thwarted.</p>
<p>Real-world and cultural sources of variability may also interfere with the quality of recordings, making certain types of measurements difficult to obtain. As in many Roma communities, living conditions do not allow for quiet recordings. According to FRA and UNDP (<xref ref-type="bibr" rid="B44">2012</xref>), the average number of persons per room is 2.7 for the Greek Roma (as opposed to just over 1 for the non-Roma population), while 35% of Greek Roma live in households without basic amenities, such as electricity, an indoor kitchen, bathroom, or toilet. Such conditions mean that quiet indoor recordings are rarely possible; recordings are likely to take place outdoors or involve bystanders. As the Roma communities discussed here are &#8220;high involvement&#8221; (<xref ref-type="bibr" rid="B111">Tannen, 1987</xref>), overlaps in conversation and multiple conversations taking place at the same time are also common. Finally, relying on spontaneous conversations also means that recordings tend to be uneven in speaking rate, volume, and pitch level and span, especially when the conversations become animated. As an indication, pitch in the speech of several female participants well exceeded 500 Hz when they were engaged in spontaneous conversation; the same women had substantially lower maxima in the semi-controlled data from QUIS, rarely exceeding 300 Hz. Similarly, the male speaker who participated in several tasks, reached a maximum of 280 Hz and rarely fell below 100 Hz when telling a story, but kept to a low level and small span of between 80 Hz and 180 Hz when taking part in various QUIS tasks.</p>
<p>In terms of analysis, the difficulties presented by the variability in the data are compounded by the fact that there is little research on Romani prosody on which to build an analysis: the bibliography of Romani linguistics by Bakker and Matras (<xref ref-type="bibr" rid="B17">2003</xref>) has more than 2,500 entries but just six publications that touch on intonation. They all treat Eastern European varieties but none that is dialectally close to Greek Thrace Romani. Thus, building on previous descriptions, as suggested by Jun &amp; Fletcher (<xref ref-type="bibr" rid="B68">2014</xref>), is not possible in this instance. As with other previously undescribed languages, an analysis can only be based on (i) general assumptions related to typology, (ii) existing knowledge of realizational variability, (iii) a finite dataset, and (iv) knowledge of neighbouring systems (though, as <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>, point out, neighbouring systems may not necessarily share typological similarities with the system under analysis).</p>
<p>Taken all together, the elements discussed above mean that any corpus of Romani is likely to include multiple sources of variability and noise both literally and figuratively: the speakers do not speak a uniform variety and code-mix using three languages, the main one of which, Romani, shows extensive dialectal variation even among small groups like those examined here. Further, the speakers are unlikely to be educated to a degree that would allow them to read aloud with ease scripted materials, certainly not in Romani, and may approach some tasks (e.g., those involving role-play) with misgivings due to their unfamiliarity. Recordings are likely to be noisy because privacy and quiet spaces are hard to come by, while conversations are animated and involve multiple participants; at the same time, community members cannot afford to travel to studios and may even be skeptical of such endeavours.</p>
</sec>
<sec>
<title>3 Data and principles of analysis</title>
<sec>
<title>3.1 The Romani data</title>
<p>The extensive variability created by the sources discussed above means that standard paradigms for data collection, even adapted to a fieldwork situation with an unschooled population, would be unlikely to be successful or would lead to a small sample in terms of number of speakers, a highly inadvisable outcome given the interspeaker variability present in the language. The Romani corpus discussed here contains instead mostly spontaneous speech from 10 speakers and a variety of speaking styles.<xref ref-type="fn" rid="n3">3</xref> Specifically, the data include the following: story-telling from three speakers, two male and one female; spontaneous conversations, involving a total of nine speakers (eight female); semi-controlled data elicited from two female and three male speakers using QUIS (<xref ref-type="bibr" rid="B107">Skopeteas et al., 2006</xref>); elicitation of words and short phrases based on the Intercontinental Dictionary Series (Ritchie Key &amp; Comrie; <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://lingweb.eva.mpg.de/ids/">http://lingweb.eva.mpg.de/ids/</ext-link>) produced by one male speaker (who also contributed one story, and took part in the QUIS tasks and in spontaneous conversations). The majority of the speakers had little or no schooling. The two youngest participants were 16 years old and the oldest was in her 50s, but most participants were in their 20s or 30s. All gave oral informed consent.</p>
</sec>
<sec>
<title>3.2 Basic principles of analysis</title>
<p>In order to analyze the Romani data the following principles were adhered to. First, the aim was to arrive at a <italic>phonological analysis</italic>, not to develop a <italic>phonetic transcription</italic> of the Romani intonational system. This aim is not specific to Romani, to this particular project, or to non-standardized linguistic varieties. Rather, it is in line with the principles underlying the development of AM: what are often referred to as <italic>AM transcriptions</italic> are in fact meant to be phonological representations characterized by underspecification (<xref ref-type="bibr" rid="B7">Arvaniti, 2011</xref>; <xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>; <xref ref-type="bibr" rid="B24">Beckman et al., 2005</xref>). This understanding of AM is in line with a more general understanding of the organization of sound systems which recognizes both the need of abstraction and the need for phonetic detail (<xref ref-type="bibr" rid="B25">Beckman et al., 2007</xref>; <xref ref-type="bibr" rid="B78">Ladd, 2011</xref>; <xref ref-type="bibr" rid="B93">Pierrehumbert, 2002</xref>).</p>
<p>This understanding of the nature and purpose of phonological representations has several consequences for analysis. First, it means that the aim of the presentations was not to more or less faithfully depict the course of <italic>F<sub>0</sub></italic>. As will be argued in more detail in Section 5.2., the course of <italic>F<sub>0</sub></italic> can be represented much more accurately by the pitch tracks themselves. Instead, the aim of the analysis here was to determine the intonational elements that are contrastive in the Romani system. Adopting this view entails that meaning cannot be dismissed and phonetic form cannot be considered without reference to meaning. Rather, the analysis follows similar lines to those used to establish the segmental contrasts in a sound system: intonational events are identified and examined in terms of their pragmatic meaning to determine whether they are contrastive in the system under analysis; meaning in this instance would involve the role played by different intonational elements in discourse (e.g., highlighting, showing finality; cf. <xref ref-type="bibr" rid="B95">Pierrehumbert &amp; Hirschberg, 1990</xref>). Decisions about the representation of the intonational elements deemed to be contrastive are based on (i) standard practice, (ii) system internal considerations (cf. <xref ref-type="bibr" rid="B54">Gussenhoven, 2007</xref>), and (iii) acceptance of what Arvaniti and Ladd (<xref ref-type="bibr" rid="B12">2009, p. 63</xref>) have called &#8220;lawful variability&#8221; (cf. <xref ref-type="bibr" rid="B33">Cangemi &amp; Grice, 2016</xref>; <xref ref-type="bibr" rid="B36">Cole &amp; Shattuck-Hufnagel, 2016</xref>; <xref ref-type="bibr" rid="B45">Frota, 2016</xref>). Each of these elements is discussed in detail below.</p>
<p>Standard practice was taken into consideration in determining the appropriate representation of tonal events. Thus, H was used to represent tones deemed to be high in a melody with respect to the speaker&#8217;s range and other tones in the same contour; L was used for tones deemed to be low by the same criteria (cf. <xref ref-type="bibr" rid="B92">Pierrehumbert, 1980, pp. 68&#8211;75</xref>). Conventions that are gradually becoming established in the field were also followed, such as Jun &amp; Fletcher&#8217;s (<xref ref-type="bibr" rid="B68">2014</xref>) recommendation to dispense with the + sign in bitonal pitch accents unless there is evidence that the two tones align independently of each other; thus here LH* is used instead of L+H*.</p>
<p>System-internal considerations mean that phonetic detail was not part of the representations unless there was evidence it was contrastive. Decisions on contrastiveness were guided by meaning in combination with form: differences in form were considered contrastive after taking into account focus and information structure and the pragmatic function of utterances in discourse (cf. <xref ref-type="bibr" rid="B92">Pierrehumbert, 1980, pp. 59&#8211;63</xref>). The analysis involved several iterations, leading to both bottom-up and top-down decisions: a first set of data determined the original analysis, which was then used to annotate more data; analytical decisions were further tested with semi-controlled QUIS data which in turn made it clear that additional refinements and revisions were necessary.</p>
<p>In addition, the analysis was kept as simple as possible until this proved untenable. In other words, rather than annotating phonetic detail and determining at a later stage if doing so was justified (as <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>, recommend), the analysis started with the simplest possible annotation labels. For instance, instead of marking rising pitch accents with both a L and a H tone (i.e., as L*+H, L+H*, L*H, LH*, etc.), only H* was originally used. Once additional data indicated that narrow focus is signalled by the use of a rising accent with consistently different realization, a distinction between H* and LH* was adopted (see Section 4.2.1.).</p>
<p>The fact that simple representations were adopted means that not all tonal events are represented in the most phonetically exhaustive way possible. For instance, in Romani H* may show a rise to a peak. This rise is an optional element determined by context and thus considered part of the accent&#8217;s phonetic realization &#8212; more specifically, of the scope of the accent&#8217;s variability &#8212; but is not seen here as essential for its representation. Although others have argued against the loss of phonetic transparency in such cases (e.g., <xref ref-type="bibr" rid="B80">Ladd &amp; Schepman, 2003</xref>), what is advocated here is standard practice in segmental phonology. For instance, in all accounts of English phonology, voiceless stops are represented as /p/, /t/ and /k/, i.e., with the IPA symbols for voiceless unaspirated plosives, even though /p<sup>h</sup>/, /t<sup>h</sup>/ and /k<sup>h</sup>/, the symbols for voiceless aspirated plosives, would provide more faithful representations. This is in line with IPA guidelines (<xref ref-type="bibr" rid="B64">IPA, 1999</xref>; for a discussion see <xref ref-type="bibr" rid="B78">Ladd, 2011</xref>); it reflects the understanding that aspiration need not be part of the symbolic representation of these phonemes since they do not contrast for aspiration with any other phonemes of English. On the other hand, in Romani, which has a three-way contrast between prevoiced, short-lag, and long-lag VOT (<xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti 2014</xref>), incorporating VOT into the phonological representation is essential. As a result of these widely accepted practices, a sound phonetically similar to English /p/ is represented as /p<sup>h</sup>/ in Romani phonology, since in that system it contrasts with unaspirated /p/. System-internal considerations comparable to those pertaining to English stops led to the decision to represent the most frequent rising accent of Romani as H* (see Section 4.2.1.).</p>
<p>As noted above, the decision to adopt representations that are as simple as possible was also based on the understanding that intonational elements exhibit lawful variability, and that such variability in intonation should be considered at least as normal as it is considered for segments (cf. <xref ref-type="bibr" rid="B36">Cole &amp; Shattuck-Hufnagel, 2016</xref>). Some of the variation observed is related to speaker and style. In the present data it was immediately evident that some participants used clearer speech than others, but also that speech clarity depended on the task: spontaneous, animated conversations showed extensive coarticulatory effects as compared to the QUIS data; these differences were evident in intonation as well.</p>
<p>Variability may also relate to dialect. This particular point could not be explored in detail here due to the extensive code-mixing of Romani (but see <xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>, for some dialectal differences in stress). Nevertheless, it is a type of variability worth discussing as it has often been neglected in intonation research. A good case in point is the contrast between H* and L+H* in English. This contrast is posited by Pierrehumbert (<xref ref-type="bibr" rid="B92">1980, ch. 4</xref>) and a pragmatic analysis of the difference between the two accents is presented in Pierrehumbert and Hirschberg (<xref ref-type="bibr" rid="B95">1990</xref>). The existence of the contrast, however, has been strongly disputed by others (see <xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 96&#8211;97</xref>, for a discussion). Indeed Ladd and Schepman (<xref ref-type="bibr" rid="B80">2003</xref>) propose that the representation (L+H)* replace H* and L+H*, on the grounds that all &#8220;sagging transitions&#8221; between high accents in English involve an <italic>F<sub>0</sub></italic> dip consistently aligned with the onset of the accented syllable. As Arvaniti and Garding (<xref ref-type="bibr" rid="B11">2007</xref>) show, however, this argument, though valid for Ladd and Schepman&#8217;s production data (which are based on one RP and one Scottish speaker), does not apply to all dialects of English. In Arvaniti and Garding&#8217;s study, speakers from Minnesota clearly followed a pattern similar to that described by Ladd and Schepman (<xref ref-type="bibr" rid="B80">2003</xref>), i.e., always used an accent which started with a clear and consistent dip and relied on pitch range to distinguish new information from contrastive focus. Southern California speakers, on the other hand, maintained an equally clear distinction between H* and L+H*, using an accent with a shallow and inconsistently present dip (H*) for new information, and an accent with a consistently present and prominent dip with stable alignment (L+H*) to indicate contrastive focus. Data like these clearly show that dialectal differences must be given due consideration in intonation research; no researcher would discuss the vowels of &#8220;English&#8221; without specifying the variety being examined, or argue about the vowel contrasts in U.S. varieties based on data from RP. The same principle should consistently apply to intonation research as well.</p>
<p>In addition to the above sources of variability, context-related lawful variation should also be considered. Some contextual factors affecting the realization of tones are discussed below. They all largely reflect aspects of tonal crowding and undershoot (<xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>; <xref ref-type="bibr" rid="B43">Fougeron &amp; Jun, 1998</xref>; <xref ref-type="bibr" rid="B48">Grabe, 1998, ch. 5</xref>; <xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 180&#8211;184</xref>; <xref ref-type="bibr" rid="B15">Arvaniti et al., 2006b</xref>).</p>
<list list-type="bullet">
<list-item>
<p>Tonal context. Tonal events are affected by proximity to other events, with tonal crowding often resulting in elision or undershoot so that pitch modulations evident in some contexts are eliminated in others; e.g., Arvaniti et al. (<xref ref-type="bibr" rid="B13">2000</xref>) show that the L tone of L*+H pitch accents in Greek can be severely undershot or eliminated altogether if L*+H accents are on adjacent syllables. Undershoot and changes in alignment are also reported with respect to the tones of the Greek <italic>wh</italic>-question melody (<xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>). The present data show undershooting of H* accents when they appear on consecutive syllables. For edge tones in particular, differences in realization depend on the location of pitch accents. As shown in detail in Section 4.2.3., L% boundary tones are manifested as low <italic>F<sub>0</sub></italic> points when the nuclear accent is in absolute phrase-final position but as low <italic>F<sub>0</sub></italic> stretches if it appears earlier. Tonal context may affect the realization of tonal events even when crowding is not an issue; cf. Venditti et al. (<xref ref-type="bibr" rid="B115">2008</xref>), who illustrate significant variation in the scaling and alignment of the Japanese accentual H*+L depending on the nature of the boundary tones that follow.</p>
</list-item>
<list-item>
<p>Location of the tone within the utterance. The prosodic position of a tonal event can also result in different realizations. Romani, for instance, shows positional variants of the H* pitch accent which is realized as a rise with peak delay in utterance-initial position but as a high fall in utterance-final position (see Section 4.2.1).</p>
</list-item>
<list-item>
<p>Interactions of stress with phrasing. Variation often depends not only on the position of a tonal event within an utterance (e.g., initial, medial, or final) but also on its precise location. In the case of pitch accents, this is determined by the position of the stressed syllables. Thus in MAE-ToBI the contrast between H* and L+H* is considered to be neutralized in absolute utterance-initial position, as the L tone of L+H* is not realized in this context (<xref ref-type="bibr" rid="B31">Brugos et al., 2006, ch. 2.5</xref>). A similar situation is observed in Greek <italic>wh</italic>-questions, which show a rise from a low point when the stressed syllable of the <italic>wh</italic>-word is not utterance-initial; the rise is truncated if the <italic>wh</italic>-word starts with the stressed syllable (<xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>; <xref ref-type="bibr" rid="B10">Arvaniti et al., 2014</xref>). In Romani, the first H* accent in an utterance is likely to show truncation in absolutely initial position as compared to its realization on a later syllable (see 4.2.1.).</p>
</list-item>
<list-item>
<p>Segmental context. Segmental context also affects the realization of tones. The presence of voiceless obstruents may obscure glissandos, a phenomenon interpreted as truncation (but see <xref ref-type="bibr" rid="B88">Niebuhr, 2008</xref>, <xref ref-type="bibr" rid="B89">2012</xref>). In the present corpus, this is evident in the realization of the H* accent which rarely shows a rise if the accented syllable starts with a voiceless obstruent. Languages may differ in how such environments are treated: Grabe (<xref ref-type="bibr" rid="B48">1998, ch. 5</xref>) reports that German shows truncation in these circumstances, while English prefers compression. Further, segmental context effects may overlap with location effects. In the Romani corpus all <italic>wh</italic>-questions started with a <italic>wh</italic>-word with initial stress and a voiceless initial consonant, such as /so/ &#8216;what&#8217;, /kon/ &#8216;who&#8217;, and /&#712;kaste/ &#8216;to whom&#8217;. Until additional evidence is available, it is assumed here that in Romani the contrast between H* and LH* is neutralized in this context. In such instances, positing the simpler representation (here H*) was preferred.</p>
</list-item>
<list-item>
<p>Speaking rate effects. Changes in speaking rate can lead to the reorganization of speech, and intonation is no exception. Fougeron and Jun (<xref ref-type="bibr" rid="B43">1998</xref>) show that in French changes in speaking rate can affect pitch range and lead to the deletion or undershoot of underlying tones. Arvaniti and Garding (<xref ref-type="bibr" rid="B11">2007</xref>) and Mixdorff et al. (<xref ref-type="bibr" rid="B87">2014</xref>) report similar patterns for English and German, respectively. In the present corpus fast, less careful speech was characterized by a greater degree of tonal undershoot and anticipatory coarticulation of tones than more careful, deliberate styles. Since spontaneous speech tends to be fast, especially when speakers are animated, effects of speaking rate must be carefully considered when determining the tonal inventory.</p>
</list-item>
<list-item>
<p>Language (and melody) specificity in choice of strategies. Examples of compression and truncation like those discussed above have led to suggestions that languages either compress or truncate (<xref ref-type="bibr" rid="B48">Grabe, 1998, ch. 5</xref>). The situation, however, is clearly more complicated than an <italic>either/or</italic> choice suggests (<xref ref-type="bibr" rid="B76">Ladd, 2008a, p. 182</xref>). In some languages at least, preferences in realization may differ depending on context. In Greek, <italic>wh</italic>-questions and consecutive L*+H accents show truncation of the L tone, as noted, but in polar questions compression is preferred for the L+H-L% edge tone configuration (<xref ref-type="bibr" rid="B15">Arvaniti et al., 2006b</xref>). Similarly, the routine calling melody of Polish shows compression of the initial rise, while the melody used for urgently calling someone shows truncation of a similar rise (<xref ref-type="bibr" rid="B16">Arvaniti et al., 2016</xref>). Given the above, it is important during analysis to keep in mind that both options, truncation and compression, may be available to speakers. This is illustrated in the Romani data as well; while many tones are eliminated, the L* accent of polar questions shows evidence of compression instead (see Section 4.2.2).</p>
</list-item>
<list-item>
<p>The nature of tones. Differences between L and H tones were discussed in Pierrehumbert (<xref ref-type="bibr" rid="B92">1980, pp. 68-75</xref>) and have been observed in several studies since (e.g., <xref ref-type="bibr" rid="B98">Prieto, 1998</xref>, <xref ref-type="bibr" rid="B99">2006</xref>). Ladd (<xref ref-type="bibr" rid="B76">2008a: 182</xref>) mentions that L tones tend to be undershot or truncated more often than H tones. Arvaniti and Garding (<xref ref-type="bibr" rid="B11">2007, p. 569</xref>) also note that L tones show more consistent alignment than H tones, &#8220;the alignment of which appears to be affected by various parameters, such as emphasis [&#8230;], metrical factors, and speaking rate.&#8221; Taken together these observations suggest that H tones may show more variable alignment, while L tones show more variable scaling. The Romani data support both observations indicating that it is important to consider whether one is dealing with a L or H tone when assessing variability.</p>
</list-item>
</list>
<p>In addition to the above, recent evidence indicates that tonal events may be manifested by a variety of means, including not just <italic>F<sub>0</sub></italic> changes but also differences in the duration, amplitude, or quality of the segments involved (e.g., <xref ref-type="bibr" rid="B16">Arvaniti et al., 2016</xref>; <xref ref-type="bibr" rid="B88">Niebuhr, 2008</xref>, <xref ref-type="bibr" rid="B89">2012</xref>; see also <xref ref-type="bibr" rid="B36">Cole &amp; Shattuck-Hufnagel, 2016</xref>, and references therein). This in turn suggests that some cues to tonal events may be redundant and thus not present at all times. For example, the LH* accent of Romani used to mark narrow focus is typically realized with a rise from a low <italic>F<sub>0</sub></italic> point and a peak within the accented vowel (see Section 4.2.1). At the same time, however, syllables associated with a LH* accent are typically longer and louder, cues that in context can be sufficient for the correct identification of the accent even if it lacks the rise from a low point or shows peak alignment later than expected. Though this is a topic that requires much work, it is worth bearing in mind when considering variability that not all instances of every tonal event will exhibit all possible traits associated with that event and that sometimes non-<italic>F<sub>0</sub></italic> cues may be the only ones present.</p>
<p>To sum up, the positing of phonological contrasts in the Romani intonational system was based on basic principles of phonological analysis, the consideration of both meaning and form, and the expectation that the realization of the posited contrastive elements would show lawful variability. Linguistic sources of variability taken into account included the phrasal position of tonal events, their interaction with the segmental, tonal, and metrical context, and language-specific preferences in resolving tonal crowding. Differences between L and H tones, the possibility of redundant cues, and speaker- and style-specific differences were also considered. Crucially, the weight attributed to these factors hinged on the role of meaning. Tonal elements were posited as contrastive if differences in form were shown to operate in discourse in a way that reflected pragmatic differences, such as the presence of focus, distinctions between given and new information, and the pragmatic function of utterances in discourse (cf. <xref ref-type="bibr" rid="B7">Arvaniti, 2011</xref>; <xref ref-type="bibr" rid="B95">Pierrehumbert &amp; Hirschberg, 1990</xref>). Given that the data were either part of QUIS, which was specifically designed to probe matters of information structure, or came from natural conversations and story-telling, in which it was possible to establish pragmatic meaning from the context and interlocutors&#8217; reactions, this practice served analysis well. Analytical decisions were revised in light of new data, particularly the semi-controlled data from QUIS, and were verified again by examining whether they remained adequate when additional spontaneous data were considered.</p>
</sec>
</sec>
<sec>
<title>4 Illustrations</title>
<sec>
<title>4.1 Stress</title>
<p>A first step to any analysis is to determine the prosodic type of the system under examination. Existing analyses of the same or related varieties can be of help. However, previous analyses should not be the only source of information, as even varieties of the same language may belong to different prosodic types (cf. <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>): e.g., Gussenhoven (<xref ref-type="bibr" rid="B52">2004, pp. 228&#8211;252</xref>) discusses varieties of Central Franconian which encode a tonal contrast absent in mainstream varieties of the German-Dutch dialectal continuum; Hualde et al. (<xref ref-type="bibr" rid="B61">2002</xref>) and Kim and Jun (<xref ref-type="bibr" rid="B70">2009</xref>) also describe varieties of Basque and Korean, respectively, that are tonal unlike the standard varieties of these languages.</p>
<p>Existing analyses report that Romani is a language with fixed stress on the ultima (<xref ref-type="bibr" rid="B85">Matras, 2002</xref>). Auditorily, this appears to apply to most of the Thrace Romani vocabulary as well. The present corpus further suggests that stressed vowels are longer and louder than unstressed vowels though quality differences are small (<xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>). Loans, however, have introduced variation in stress location. For instance, when words are borrowed from Turkish, they often acquire Romani morphology; the addition of suffixes in particular leads to stress shifting to the penult; e.g., a Turkish word like <italic>pembe</italic> &#8216;pink&#8217; when used with a feminine noun acquires a feminine suffix <italic>&#8211;a</italic> yielding /pem&#712;bea/ &#8216;pink.F&#8217; with penultimate stress. Such examples are not uncommon (<xref ref-type="bibr" rid="B2">Adamou &amp; Arvaniti, 2014</xref>).</p>
<p>Crucially, in Romani declaratives with broad focus all words show a pitch rise or high pitch on the syllable perceived as stressed; this applies whether the utterance presents new or given information. This is illustrated in Figure <xref ref-type="fig" rid="F2">2</xref>: as can be seen, even function words like the preposition [kaj] and the classifier [ta&#712;neja] show a pitch rise; such pitch rises on function words are a common occurrence in the corpus. Data like these could lead to the conclusion that this variety of Romani has a lexical pitch accent system in which one syllable per word carries a rising melody, or that high or rising <italic>F<sub>0</sub></italic> is a feature of stress.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of &#8220;all new&#8221; utterance from QUIS, illustrating the use of accentuation even on function words such as [kaj] &#8216;at&#8217; and [ta&#712;neja], a classifier. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav2">http://dx.doi.org/10.5334/labphon.14.wav2</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74465/"/>
</fig>
<p>The connection between stress and high or rising pitch is still well accepted thanks to early work on the topic (<xref ref-type="bibr" rid="B47">Fry, 1958</xref>) and despite plenty of subsequent research clarifying the relationship between stress and intonation (e.g., <xref ref-type="bibr" rid="B22">Beckman, 1986</xref>; <xref ref-type="bibr" rid="B23">Beckman &amp; Edwards, 1994</xref>). Here the standard view of the autosegmental-metrical (AM) framework of intonational phonology is adopted, namely that stress is independent of changes in pitch related to intonation (<xref ref-type="bibr" rid="B7">Arvaniti, 2011</xref>; <xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 49&#8211;55</xref>). The connection between the two is indirect: stress is determined by metrical structure; in turn, stressed (metrically prominent) syllables are licensed for association with a pitch accent but need not always be accented.</p>
<p>Data like those in Figure <xref ref-type="fig" rid="F2">2</xref> demonstrate why it is crucial to examine not only declaratives that present new information, as is customarily the case, but other types of utterances as well, including questions and declaratives with early narrow focus (cf. <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>). Doing so allows us to separate the effects of stress from those of intonation. Evidence from questions and narrow focus utterances makes it clear that pitch rises in Romani are not an exponent of stress or a lexical property of words, but an independent phenomenon. Sentences with early narrow focus show that pitch rises are not present postfocally. This is seen in Figure <xref ref-type="fig" rid="F3">3</xref> in which only the negative particle [naj] is accented, while content words [maj&#712;muna] &#8216;monkey&#8217; and [a&#712;ia] &#8216;bear&#8217; show falling and flat <italic>F<sub>0</sub></italic>, respectively. The same applies to <italic>wh</italic>-questions, like that in Figure <xref ref-type="fig" rid="F4">4</xref>: there is only one marked pitch movement, that on the <italic>wh</italic>-word [so] &#8216;what&#8217;, after which <italic>F<sub>0</sub></italic> drops until the end of the utterance. Utterances like these clearly show that typologically Thrace Romani is a linguistic variety which has stress and uses pitch primarily to encode intonational differences.</p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from a QUIS game, illustrating the use of low (flat or falling) <italic>F<sub>0</sub></italic> on stressed syllables. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav3">http://dx.doi.org/10.5334/labphon.14.wav3</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74466/"/>
</fig>
<fig id="F4">
<label>Figure 4</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of a <italic>wh</italic>-question from spontaneous conversation, illustrating the lack of <italic>F<sub>0</sub></italic> rises on content words after the <italic>wh</italic>-word [so] &#8216;what&#8217;. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav4">http://dx.doi.org/10.5334/labphon.14.wav4</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74467/"/>
</fig>
</sec>
<sec>
<title>4.2 Tonal inventory</title>
<p>The above discussion of stress indicates that pitch modulation in Romani should be treated as postlexical, i.e., as intonation. The next step then is to determine the number and nature of tonal events &#8212; pitch accents and edge tones &#8212; and their use in the system.</p>
<sec>
<title>4.2.1 High pitch accents</title>
<p>The discussion of stress clearly showed that stressed syllables are often, though not always, realized with rising or high pitch. Rising and high pitch are interpreted here as reflexes of a H* pitch accent. The corpus indicates that the H* accent can take several forms: sometimes it shows a rise from a low point, while at other times it is manifested as high <italic>F<sub>0</sub></italic>, a plateau, or a fall. These different realizations can be seen in Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F4">4</xref>, <xref ref-type="fig" rid="F5">5</xref>, <xref ref-type="fig" rid="F6">6</xref>, <xref ref-type="fig" rid="F7">7</xref>, <xref ref-type="fig" rid="F8">8</xref>, <xref ref-type="fig" rid="F10">10</xref>, <xref ref-type="fig" rid="F11">11</xref>, <xref ref-type="fig" rid="F12">12</xref>, <xref ref-type="fig" rid="F14">14</xref>, <xref ref-type="fig" rid="F16">16</xref>, and <xref ref-type="fig" rid="F17">17</xref>. Most of the observed variation in the realization of H* can be explained by context. The first accent in an utterance is usually realized with a substantial rise and delayed peak (see Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F5">5</xref>, and <xref ref-type="fig" rid="F8">8</xref>). Most subsequent accents do not exhibit either of these characteristics except in careful speech: compare Figure <xref ref-type="fig" rid="F2">2</xref> from QUIS and Figure <xref ref-type="fig" rid="F5">5</xref> from a spontaneous and rather animated conversation. On the other hand, the accentual rise may be barely present if the utterance starts with a voiceless consonant; this is shown in Figure <xref ref-type="fig" rid="F4">4</xref> in which the H* accent is on [so] &#8216;what&#8217;.</p>
<fig id="F5">
<label>Figure 5</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of a hypothetical followed by a <italic>wh</italic>-question; data from spontaneous conversation, illustrating the variable realization of the H* pitch accent. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav5">http://dx.doi.org/10.5334/labphon.14.wav5</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74468/"/>
</fig>
<fig id="F6">
<label>Figure 6</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of a broad focus utterance from QUIS, illustrating the realization of H* in different contexts, including in absolutely utterance-final position. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav6">http://dx.doi.org/10.5334/labphon.14.wav6</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74469/"/>
</fig>
<fig id="F7">
<label>Figure 7</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of a broad focus utterance from spontaneous conversation, illustrating the realization of H* in different contexts, including in nuclear but not absolutely utterance-final position. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav7">http://dx.doi.org/10.5334/labphon.14.wav7</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74470/"/>
</fig>
<fig id="F8">
<label>Figure 8</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from a QUIS map-task with narrow contrastive focus on the final word. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav8">http://dx.doi.org/10.5334/labphon.14.wav8</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74471/"/>
</fig>
<p>Unlike prenuclear H* accents, which often show a rise, utterance-final words typically show a fall that starts on the stressed syllable. This is illustrated in Figure <xref ref-type="fig" rid="F6">6</xref> (a similar final accent can be seen in Figure <xref ref-type="fig" rid="F10">10</xref> on [merde&#712;fea] &#8216;ladder&#8217;). Figure <xref ref-type="fig" rid="F6">6</xref> includes three accents: the first is realized as a rise that spans the entire accented vowel, the second as high <italic>F<sub>0</sub></italic>, and the last one as a fall throughout the stressed vowel of [lo&#712;le] &#8216;red&#8217;. The difference appears to be context-related with the final accent being realized as a fall under pressure from the upcoming L% (on edge tones; see Section 4.2.3.). On the other hand, when tonal crowding is reduced, as in Figure <xref ref-type="fig" rid="F7">7</xref> where the last word has antepenultimate stress, the H* accent may be realized as high <italic>F<sub>0</sub></italic> instead. Given that the differences in realization can be explained by context and the location of stress, they do not warrant a phonological distinction: there is no evidence that final accents in sentences like those in Figures <xref ref-type="fig" rid="F5">5</xref>, <xref ref-type="fig" rid="F6">6</xref>, <xref ref-type="fig" rid="F7">7</xref>, or <xref ref-type="fig" rid="F10">10</xref> serve any different purpose than prenuclear accents. As noted earlier, in sentences encoding &#8220;all new&#8221; or given information, all words are accented, suggesting that the main function of the accents is to highlight stressed syllables (<xref ref-type="bibr" rid="B8">Arvaniti &amp; Adamou, 2011</xref>; cf. <xref ref-type="bibr" rid="B32">Calhoun, 2010</xref>). A non-exhaustive presentation of the variation of H* is given in Table <xref ref-type="table" rid="T1">1</xref>.</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>Schematized <italic>F<sub>0</sub></italic> contour (continuous line) depicting context-dependent realizations of the H* pitch accent on the target syllable (grey box) and, where applicable, on neighboring syllables (white boxes). This is not an exhaustive list of possible H* realizations in Romani.</p>
</caption>
<table>
<tr>
<th valign="top" align="left">Context</th>
<th valign="top" align="left">Realization</th>
<th valign="top" align="left">Illustration</th>
</tr>
<tr>
<td colspan="3">
<hr/></td>
</tr>
<tr>
<td valign="top" align="left">Utterance initially</td>
<td valign="top" align="left"></td>
<td valign="top" align="left"></td>
</tr>
<tr>
<td valign="top" align="left">On first syllable</td>
<td valign="top" align="left"><inline-graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74481/"/></td>
<td valign="top" align="left">Figures <xref ref-type="fig" rid="F4">4</xref>, <xref ref-type="fig" rid="F6">6</xref></td>
</tr>
<tr>
<td valign="top" align="left">On non-initial syllable</td>
<td valign="top" align="left"><inline-graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74482/"/></td>
<td valign="top" align="left">Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F5">5</xref>, <xref ref-type="fig" rid="F7">7</xref></td>
</tr>
<tr>
<td valign="top" align="left">Utterance medially</td>
<td valign="top" align="left"></td>
<td valign="top" align="left"></td>
</tr>
<tr>
<td valign="top" align="left">Adjacent to other accent</td>
<td valign="top" align="left"><inline-graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74483/"/></td>
<td valign="top" align="left">Figure <xref ref-type="fig" rid="F2">2</xref></td>
</tr>
<tr>
<td valign="top" align="left">Non-adjacent to other accent</td>
<td valign="top" align="left"><inline-graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74484/"/></td>
<td valign="top" align="left">Figures <xref ref-type="fig" rid="F5">5</xref>, <xref ref-type="fig" rid="F6">6</xref>, <xref ref-type="fig" rid="F7">7</xref></td>
</tr>
<tr>
<td valign="top" align="left">Utterance finally</td>
<td valign="top" align="left"></td>
<td valign="top" align="left"></td>
</tr>
<tr>
<td valign="top" align="left">Final stress</td>
<td valign="top" align="left"><inline-graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74485/"/></td>
<td valign="top" align="left">Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F6">6</xref></td>
</tr>
<tr>
<td valign="top" align="left">Non-final stress</td>
<td valign="top" align="left"><inline-graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74486/"/></td>
<td valign="top" align="left">Figures <xref ref-type="fig" rid="F7">7</xref>, <xref ref-type="fig" rid="F10">10</xref></td>
</tr>
</table>
</table-wrap>
<p>There are, however, realizations of high accents in Romani which indicate that not all can be represented as H*. In utterance-final position, one can observe a difference between accents in broad focus utterances, which show the flat or falling <italic>F<sub>0</sub></italic> discussed above, and accents with a marked dip and a rise-fall contour as on the word [a&#712;ra&#967;ni] &#8216;spider&#8217; in Figure <xref ref-type="fig" rid="F8">8</xref>. This accent is represented as LH*. The two accent types serve distinct purposes: LH* is used to mark narrow focus in declaratives (cf. <xref ref-type="bibr" rid="B95">Pierrehumbert &amp; Hirschberg, 1990</xref>). The same accent is also found in early narrow focus, as in Figure <xref ref-type="fig" rid="F9">9</xref>: here <italic>F<sub>0</sub></italic> lowers from the onset of the second phrase to the onset of the stressed syllable of [t&#643;ala&#712;vel] before rising to a peak within this syllable; following words are unaccented (<xref ref-type="bibr" rid="B8">Arvaniti &amp; Adamou, 2011</xref>). The realization of this accent can be juxtaposed to the H* accent on [fu&#712;lel] &#8216;descend&#8217; in Figure <xref ref-type="fig" rid="F10">10</xref> which encodes new information: though this accent also shows a significant pitch rise (being phrase-initial), its range is reduced relative to LH* (the speaker is the same in Figures <xref ref-type="fig" rid="F9">9</xref> and <xref ref-type="fig" rid="F10">10</xref>); the peak is aligned with the end of the accented syllable, while the following noun [merde&#712;fea] &#8216;ladder&#8217; is also accented. The difference between LH* and H* in focal position can also be seen in Figure <xref ref-type="fig" rid="F11">11</xref> which includes a short exchange during a QUIS game revolving around the word [a&#712;ia] &#8216;bear&#8217;: the first token is contrastive and bears a LH* accent, while the second is the interlocutor&#8217;s confirmation and thus given information and bearing H* instead; a difference in overall scaling between the two accents, in addition to shape, is obvious here too. Given the above, the difference between H* and LH* cannot be attributed to a simple expansion of pitch range as has been advocated by Ladd for English (e.g., <xref ref-type="bibr" rid="B79">Ladd &amp; Morton, 1997</xref>); pitch level and span are both involved, but there are also differences in alignment. Further, the low <italic>F<sub>0</sub></italic> at the onset of the accented syllable is systematic for LH*, and can be the outcome of a drop in <italic>F<sub>0</sub></italic> (Figure <xref ref-type="fig" rid="F8">8</xref>), a low stretch (Figure <xref ref-type="fig" rid="F9">9</xref>), or a combination of the two depending on context. The rise, especially in final position in a long utterance, can be small: it is just sufficient for the accent to sound high rather than falling and in this position it gives the accent its characteristic rise-fall shape (Figures <xref ref-type="fig" rid="F8">8</xref> and <xref ref-type="fig" rid="F10">10</xref>). In short, an analysis with only one accent, whether this is represented as H* or LH*, does not appear to be satisfactory for Romani.</p>
<fig id="F9">
<label>Figure 9</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from QUIS elicitation with narrow contrastive focus on the verb. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav9">http://dx.doi.org/10.5334/labphon.14.wav9</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74472/"/>
</fig>
<fig id="F10">
<label>Figure 10</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of a broad focus utterance from QUIS elicitation encoding &#8220;all new&#8221; information. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav10">http://dx.doi.org/10.5334/labphon.14.wav10</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74473/"/>
</fig>
<fig id="F11">
<label>Figure 11</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from a QUIS game illustrating the differences between LH*, H*, and L* accents on the same word [a&#712;ia] &#8216;bear&#8217;. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav11">http://dx.doi.org/10.5334/labphon.14.wav11</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74474/"/>
</fig>
<p>The above data beg the question: why not posit instead that Romani has a H*L accent in utterance final position and a LH* accent elsewhere, realized with an optional L component that is present particularly when the accent is used in corrective or contrastive contexts? This analysis would be phonetically transparent and faithful to the most frequent realizations of the two accents. The reason why such a solution is not adopted has to do with the function of the accents within the Romani intonational system. If the above analysis were adopted, Romani would be said to have a LH* accent that has a variety of functions: it is used both to mark new information and for metrical purposes but can also mark narrow focus when needed. The H*L is also used to mark new information (but only at the end of utterances) and can also serve metrical purposes. Relying on the role of the accents in the system makes it clear that this analysis is not optimal as it posits two accents on the basis of form but with mixed functions. Further, this analysis assumes that the ubiquitous <italic>F<sub>0</sub></italic> fall at the end of declaratives is due sometimes to a L edge tone (after LH*) and sometimes to the accent itself (after H*L). The consequences of particular pitch accent choices for the analysis of edge tones are discussed in more detail in Section 4.2.3.</p>
</sec>
<sec>
<title>4.2.2 Low pitch accents</title>
<p>Romani also shows low pitch accents. These appear in two main environments in the corpus: before a continuation rise and in polar questions. Low accents in both cases are represented as L* and serve to highlight the word in focus in environments indicating non-completion.</p>
<p>A canonical instantiation of L* is seen in Figure <xref ref-type="fig" rid="F12">12</xref> on [&#712;mat&#643;ka] &#8216;cat&#8217; which has low, flat <italic>F<sub>0</sub></italic> on its stressed syllable. Figure <xref ref-type="fig" rid="F13">13</xref> exemplifies the realization of L* in a polar question. Like the H* accent, L* exhibits realization variability. In general, L* is realized as a dip or low-<italic>F<sub>0</sub></italic> stretch that is more pronounced and longer compared to the dip of rising accents discussed in Section 4.2.1.; as a result, the accent sounds low not high. The difference in the extent of the low <italic>F<sub>0</sub></italic> stretch is illustrated in Figure <xref ref-type="fig" rid="F11">11</xref> which includes a LH*, a H* and a L* accent on the word [a&#712;ia] &#8216;bear&#8217;. Low <italic>F<sub>0</sub></italic>, however, is often realized on the syllable <italic>preceding</italic> the one with stress, while the stressed syllable itself is low but rising. This happens particularly if the stressed syllable is phrase-final; this can be observed in the first phrase in Figures <xref ref-type="fig" rid="F9">9</xref> and <xref ref-type="fig" rid="F10">10</xref>, where [mu&#712;ru&#643;] &#8216;man&#8217; and [t&#643;&#688;o&#712;ri] &#8216;girl&#8217;, respectively, show a deliberate dip on their first (unstressed) syllable. It is only when the stressed syllable is further from the boundary tone that the L* is fully realized, as in Figures <xref ref-type="fig" rid="F11">11</xref> and <xref ref-type="fig" rid="F12">12</xref>. However, the preponderance of final stress means that such realizations of L* are not very frequent.</p>
<fig id="F12">
<label>Figure 12</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from a QUIS game illustrating the realization of L* in the absence of tonal crowding. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav12">http://dx.doi.org/10.5334/labphon.14.wav12</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74475/"/>
</fig>
<fig id="F13">
<label>Figure 13</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from a QUIS game illustrating the polar question melody of Romani when no tonal crowding is involved. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav13">http://dx.doi.org/10.5334/labphon.14.wav13</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74476/"/>
</fig>
<p>The dip in <italic>F<sub>0</sub></italic> reflecting a L* pitch accent can be less pronounced in the case of polar questions; e.g., in Figure <xref ref-type="fig" rid="F13">13</xref>, [i&#712;klan] starts low but <italic>F<sub>0</sub></italic> rises smoothly afterwards. The fact that the dip in questions like that in Figure <xref ref-type="fig" rid="F13">13</xref> is the reflex of a L* is supported by utterances like that in Figure <xref ref-type="fig" rid="F14">14</xref>: the melody here shows a low <italic>F<sub>0</sub></italic> stretch on the last vowel of [la&#712;t&#643;&#688;o] &#8216;nice&#8217; which is the focus of the question and is followed by the HL% boundary tone typical of polar questions (see Section 4.2.3.). While the boundary L is undershot in this instance, due to tonal crowding, it is clear that the <italic>F<sub>0</sub></italic> dip associated with the L* focal accent is considered essential for the melody and thus fully realized by elongating the last vowel of [la&#712;t&#643;&#688;o] &#8216;nice&#8217;.</p>
<fig id="F14">
<label>Figure 14</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from spontaneous conversation illustrating the polar question melody of Romani in the presence of tonal crowding. Click on the figure to listen to the sound file. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav14">http://dx.doi.org/10.5334/labphon.14.wav14</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74477/"/>
</fig>
<p>The decision to analyze these accents as L* when most instances are characterized by undershoot due to extensive coarticulation with upcoming H tones may be met with skepticism. Could it be, for example, that questions use the same LH* accent as statements to indicate narrow focus? Why not use H* for the accent in continuation rises, since <italic>F<sub>0</sub></italic> is often rising on accented syllables? The answers to these questions lie in basic principles for distinguishing L and H accents, system internal considerations, and analytical coherence (<xref ref-type="bibr" rid="B54">Gussenhoven, 2007</xref>).</p>
<p>First, L* accents in both continuation rises and polar questions show scaling that is low relative to the speaker&#8217;s range and other accents, as shown in Figures <xref ref-type="fig" rid="F12">12</xref>, <xref ref-type="fig" rid="F13">13</xref>, <xref ref-type="fig" rid="F14">14</xref>. Second, polar questions in particular always end in a L edge tone (see Figures <xref ref-type="fig" rid="F13">13</xref> and <xref ref-type="fig" rid="F14">14</xref>); if so, then analysing their melody as LH* L% would make them identical to narrow-focused statements. This is patently false, however, and this stands to reason: speakers should wish to differentiate statements from questions. The difference has primarily to do with the shape of the pitch rise and the location of the peak. In narrow focus statements, the rise and fall are symmetrical; the rise is convex in shape and the peak is typically reached on the accented syllable, after which <italic>F<sub>0</sub></italic> begins to fall. In polar-questions, the contour starts with a low <italic>F<sub>0</sub></italic> stretch, while the rise is concave and followed by a fall of relatively short duration. This difference in shape is illustrated in Figure <xref ref-type="fig" rid="F15">15</xref> which shows the contour of the word [a&#712;ia] &#8216;bear&#8217; with LH* (from Figure <xref ref-type="fig" rid="F11">11</xref>) and as a polar question (from a different speaker with different pitch range but almost identical duration). The differences support the observation above that the H tone of the LHL sequence in polar questions occurs close to the end of the question (see Figure <xref ref-type="fig" rid="F13">13</xref>). Due to the limited variation in stress location in Romani, it is not clear if the phrase-final, phrase-penultimate, or last <italic>stressed</italic> syllable is the docking site of this H tone (though Figure <xref ref-type="fig" rid="F13">13</xref> and other similar examples suggest it is the last stressed syllable, as in Greek; <xref ref-type="bibr" rid="B15">Arvaniti et al., 2006b</xref>; <xref ref-type="bibr" rid="B50">Grice et al., 2000</xref>). Despite this uncertainty, it is clear that the H is not aligned with the accented syllable of the word in focus unless this word is phrase-final. This suggests that the H tone is less likely to be part of the pitch accent itself and more likely to be part of an edge tone. These considerations lead to the overall analysis of the polar question melody as L* HL%. The only alternative analysis involving a LH* accent would be to represent the polar question melody as LH* HL%. But this would entail the presence of a plateau between the two H tones; this is not attested, however, although plateaus are frequent in Romani (see Section 4.2.4.)</p>
<fig id="F15">
<label>Figure 15</label>
<caption>
<p><italic>F<sub>0</sub></italic> contours (in Hz) of the word [a&#712;ia] &#8216;bear&#8217; with L* HL% (gray line) and LH* L% (black line). This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav15a">http://dx.doi.org/10.5334/labphon.14.wav15a</ext-link> (gray line) and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav15b">http://dx.doi.org/10.5334/labphon.14.wav15b</ext-link> (black line).</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74478/"/>
</fig>
<p>A viable alternative would be to represent the accent of polar questions as L*H and the entire melody as L*H L%, accepting that the accentual H is aligned independently of the L* tone (see <xref ref-type="bibr" rid="B54">Gussenhoven, 2007</xref>, for arguments that pitch accent tones need not be bound to each other). If so, then the posited L* used in continuation rises could be seen as the flip side of H*, an accent used primarily for metrical purposes; as such, this accent can be elided or severely undershot. Its presence is required simply to create a perceptual contrast with the upcoming H% boundary tone (cf. <xref ref-type="bibr" rid="B54">Gussenhoven, 2007</xref>). A similar reversal of polarity is reported for Greek by Baltazani and Jun (<xref ref-type="bibr" rid="B19">1999</xref>). At present it is not possible to determine which analysis is optimal. This is due both to the fact that the corpus contains relatively few instances of polar questions and continuation rises and because we do not as yet have strict criteria in intonational research to assess alternative analyses (but see <xref ref-type="bibr" rid="B103">Ritter and Grice, 2015</xref>, and <xref ref-type="bibr" rid="B55">Gussenhoven, 2016</xref>). I return to this point in Section 5.3.</p>
</sec>
<sec>
<title>4.2.3 Edge tones and phrasing</title>
<p>Edge tones are often discussed together with phrasing. Following Pierrehumbert (<xref ref-type="bibr" rid="B92">1980</xref>), many analyses have adopted two types of edge tones, phrase accents and boundary tones (e.g., L- and L% respectively). Since Beckman and Pierrehumbert (<xref ref-type="bibr" rid="B26">1986</xref>), these two types of edge tones have been linked to distinct levels of phrasing, the intermediate phrase (ip) for phrase accents and the Intonational Phrase (IP) for boundary tones. Evidence in favour of the independence of phrase accents and boundary tones in such configurations has been reported <italic>inter alia</italic> in Arvaniti et al. (<xref ref-type="bibr" rid="B15">2006b</xref>) for the Greek L+H-, Barnes et al. (<xref ref-type="bibr" rid="B20">2006</xref>) for the English L-, and Arvaniti and Ladd (<xref ref-type="bibr" rid="B12">2009</xref>) for the Greek L-. On the other hand, the need for two levels of phrasing has been hotly disputed by some (e.g., <xref ref-type="bibr" rid="B52">Gussenhoven, 2004, pp. 316&#8211;319</xref>; <xref ref-type="bibr" rid="B74">Ladd, 1983</xref>; see <xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 142&#8211;147</xref> for a discussion). Nevertheless, combinations of tones are clearly needed for observational adequacy, independently of whether one adopts the notion of a phrase accent and relates it to the presence of two levels of phrasing. For instance, the presence of two H tones each with its own target accounts for final rises in English which show a step-up from one high pitch level to the next (<xref ref-type="bibr" rid="B31">Brugos et al., 2006</xref>; <xref ref-type="bibr" rid="B92">Pierrehumbert, 1980</xref>). Similarly, Ritchart and Arvaniti (<xref ref-type="bibr" rid="B102">2014</xref>) analyze Southern California uptalk as L* L-H%; the L-H% edge tone configuration accounts for the late onset and low scaling of the uptalk rise as compared to rises in questions, which are analyzed as L* H-H%. Independently of whether one assumes that one of these tones is a phrase accent and the other a boundary tone each demarcating a different phrasal constituent, it is clear that both are needed to adequately represent this difference between questions and statements with uptalk.</p>
<p>In order to determine whether a language has one or two levels of phrasing, Jun and Fletcher (<xref ref-type="bibr" rid="B68">2014</xref>) propose that one uses disambiguation (of the <italic>Mary is not drinking because she is unhappy</italic> type) or increasingly longer utterances in which &#8220;weight&#8221; is added to specific constituents. The assumption is that these manipulations will break down long utterances into shorter phrases. One can then examine if these shorter phrases are comparable to longer ones or present their own characteristics.<xref ref-type="fn" rid="n4">4</xref> A somewhat different approach is adopted by Arvaniti and Baltazani (<xref ref-type="bibr" rid="B9">2005</xref>) in the GRToBI analysis of the Greek intonational system. The authors annotated ips and IPs based on (impressionistic) degree of juncture, then compared the two types: they found that phrases annotated as ips had less complex tonal movements (simple rise or fall) and less extreme scaling that those annotated as IPs; e.g., while at the end of IPs Greek speakers reached the bottom of their range, they did not do so at the end of ips ending in L-.</p>
<p>The procedure of Arvaniti and Baltazani (<xref ref-type="bibr" rid="B9">2005</xref>) is not easy to use with a diverse corpus like that of Romani. Establishing a speaker&#8217;s pitch range in the laboratory is much easier than in natural speech, particularly when speakers touch upon sensitive topics, become excited, etc. (see Section 2). The approach of Jun and Fletcher (<xref ref-type="bibr" rid="B68">2014</xref>) is more appropriate in such circumstances, if suitable data are available. In the present corpus, however, attempts to elicit such longer utterances resulted in short phrases separated by prolonged pauses; Figures <xref ref-type="fig" rid="F8">8</xref>, <xref ref-type="fig" rid="F9">9</xref>, <xref ref-type="fig" rid="F10">10</xref>, and <xref ref-type="fig" rid="F16">16</xref> illustrate the types of substantial breaks speakers of Romani used when at most a minor break would be expected. This can be juxtaposed to spontaneous animated speech in which expected breaks are missing; e.g., in Figure <xref ref-type="fig" rid="F5">5</xref> there is no break between the subordinate and main clauses. Thus, although there are perceived differences in strength between some phrasal boundaries, it is not possible to discern systematic differences between them in terms of function, scaling, or tonal configuration. In turn this suggests that one level of phrasing is sufficient for Romani (barring new data). The edge tones include L%, H%, HL%, and LH%. HL% is found at the end of polar questions, as in Figures <xref ref-type="fig" rid="F13">13</xref> and <xref ref-type="fig" rid="F14">14</xref>. LH% is attested in <italic>wh</italic>-questions (not illustrated).</p>
<p>A reason why researchers posit two types of edge tones is that phrase accents often fill the gap between the last pitch accent and the end of the utterance or show secondary association (<xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>; <xref ref-type="bibr" rid="B20">Barnes et al., 2006</xref>; <xref ref-type="bibr" rid="B50">Grice et al., 2000</xref>). In Romani there is no evidence for secondary association.<xref ref-type="fn" rid="n5">5</xref> Spreading appears to apply only to the L% boundary tone which spreads to the left when focus is early: in such instances, <italic>F<sub>0</sub></italic> starts dropping towards the end of the stressed syllable of the accented word and remains low for the remainder of the utterance, though no consistent pattern for the extent of the spread can be discerned (cf. Figures <xref ref-type="fig" rid="F3">3</xref>, <xref ref-type="fig" rid="F4">4</xref>, and <xref ref-type="fig" rid="F9">9</xref>). Figure <xref ref-type="fig" rid="F7">7</xref> shows a different instance of L% spreading: here, <italic>F<sub>0</sub></italic> falls immediately after the stressed antepenult of [&#712;gomeno] &#8216;boyfriend&#8217; so that the last two syllables in the utterance are both low in pitch.</p>
<p>Finally, it is worth noting that the present analysis of edge tones follows the established practice of separating final (nuclear) pitch movements into a pitch accent and following edge tones. Thus &#8220;nuclear falls&#8221; in Romani declaratives are analyzed as a sequence of a H* pitch accent and a L% boundary tone. This type of analysis goes back to Pierrehumbert (<xref ref-type="bibr" rid="B92">1980, ch. 1</xref>) who analyzed English nuclear falls as consisting of a H* pitch accent followed by a L-L% edge tone configuration. This is not the only possibility, however. Gussenhoven (<xref ref-type="bibr" rid="B52">2004, pp. 296&#8211;299</xref>) analyses the same English nuclear fall as consisting of a H*L pitch accent followed by a L% boundary tone or L<italic><sub>&#953;</sub></italic> in his notation (for additional arguments for &#8220;off ramp&#8221; analyses of English melodies, see <xref ref-type="bibr" rid="B55">Gussenhoven, 2016</xref>; see also Peters, Hanssen &amp; Gussenhoven (<xref ref-type="bibr" rid="B91">2015</xref>) for &#8220;off ramp&#8221; analyses of a number of Germanic varieties). A discussion of the two views is beyond the scope of the paper, but it is worth keeping in mind when determining how best to analyze a particular language that any analysis of edge tones hinges on decisions about the accent inventory and vice versa.</p>
</sec>
<sec>
<title>4.2.4 The use of plateaux</title>
<p>In the Romani corpus, plateaux are quite frequent (see, e.g., Figures <xref ref-type="fig" rid="F2">2</xref>, <xref ref-type="fig" rid="F3">3</xref>, <xref ref-type="fig" rid="F6">6</xref>, <xref ref-type="fig" rid="F7">7</xref>, <xref ref-type="fig" rid="F8">8</xref>, and <xref ref-type="fig" rid="F12">12</xref>). An utterance in which plateaux are used almost exclusively is shown in Figure <xref ref-type="fig" rid="F16">16</xref>. In some languages differences between peaks and plateaux are meaningful; this applies, e.g., to Neapolitan Italian (<xref ref-type="bibr" rid="B40">D&#8217;Imperio, 2000</xref>; <xref ref-type="bibr" rid="B41">D&#8217;Imperio et al., 2000</xref>). In others, like British English, it is clear that plateaux affect estimates of pitch accent scaling but not necessarily accent identity (<xref ref-type="bibr" rid="B71">Knight, 2008</xref>). In Romani, however, peaks and plateaux appear to be realizational variants of L and H tones both phrasal and accentual, so that plateaux and glissandos are interchangeable. Compare Figures <xref ref-type="fig" rid="F16">16</xref> and <xref ref-type="fig" rid="F17">17</xref>, both showing utterances elicited from the same speaker during a QUIS task. In Figure <xref ref-type="fig" rid="F16">16</xref> the <italic>F<sub>0</sub></italic> of almost every syllable is flat, independently of stress and association with a tone (cf. unaccented [ka] in [ka&#712;fe] and accented [&#712;pa] in [&#712;pasta]). In Figure <xref ref-type="fig" rid="F17">17</xref>, on the other hand, plateaux and glissandos coexist: the H* accents are realized as rises but the two initial (unstressed) syllables and the final H% are realized as plateaux. The fact that glissandos and plateaux can be used interchangeably and mixed in the same utterance indicates that there is little difference between them in Romani. Plateaux appear to be more frequent in ritualistic and formal speech, such as story-telling and QUIS games, respectively. Realizations of H%, plateaux are frequent as a floor-holding device, particularly when pauses mid-utterance are involved, as in Figure <xref ref-type="fig" rid="F16">16</xref> (this use is akin to plateaux attested in Greek and some varieties of English; cf. <xref ref-type="bibr" rid="B9">Arvaniti &amp; Baltazani, 2005</xref>, on Standard Greek; <xref ref-type="bibr" rid="B35">Clopper &amp; Smiljanic, 2011</xref>, and <xref ref-type="bibr" rid="B102">Ritchart &amp; Arvaniti, 2014</xref>, on U.S. English varieties). Overall, these observations hint at a stylistic rather than a pragmatic difference between plateaux and glissandos in Romani, indicating that the difference need not be part of the phonological representation.</p>
<fig id="F16">
<label>Figure 16</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from QUIS, illustrating the use of plateaux. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav16">http://dx.doi.org/10.5334/labphon.14.wav16</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74479/"/>
</fig>
<fig id="F17">
<label>Figure 17</label>
<caption>
<p>Spectrogram, <italic>F<sub>0</sub></italic> contour (in Hz), AM annotation, and gloss of an utterance from QUIS, illustrating the mixing of plateaux and glissandos in the same utterance; the speaker is the same as in Figure <xref ref-type="fig" rid="F16">16</xref>. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.14.wav17">http://dx.doi.org/10.5334/labphon.14.wav17</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6174/file/74480/"/>
</fig>
</sec>
</sec>
</sec>
<sec>
<title>5 Discussion</title>
<sec>
<title>5.1 Laboratory and spontaneous data</title>
<p>The above presentation of some elements of the Romani prosodic system shows that the principles used here allow for the development of a phonological analysis even when the data present multiple sources of variation. The variety of speech styles included in the corpus allowed for a more robust analysis: variability was present and had to be taken into consideration, while decisions were not based on a uniform (and, for that reason, possibly unrepresentative) dataset as is typical of laboratory studies.</p>
<p>This does not mean that laboratory data are not useful or should be dispreferred in research. Indeed, the analysis presented here serves to show that it is counterproductive to pit laboratory and spontaneous data against each other, considering one or the other inherently superior or better suited for research (cf. <xref ref-type="bibr" rid="B117">Xu, 2010</xref>). Rather, what is proposed and illustrated here is a back and forth between the two: spontaneous data allow one to establish a set of hypotheses about the system under analysis; these can then be tested by means of controlled or semi-controlled data; any changes should be subjected to new scrutiny using spontaneous data and, if necessary, to further revision. Thus, the present work shows that it is possible to use spontaneous and (semi-)controlled data synergistically and that each type can provide answers to particular problems during analysis. Given the importance of meaning advocated here, however, approaching a previously undescribed intonational system using primarily spontaneous data was advantageous, as such data include a wealth of information in terms of both linguistic and pragmatic context that can serve as analytical tools; e.g., new, given, and contrastive information could be tracked from discourse, and linguistic context could provide clues as to the reasons for variation.</p>
</sec>
<sec>
<title>5.2 Problems with a level of broad phonetic transcription</title>
<p>The present corpus illustrates issues that can arise when variability clashes with established notions of uniformity in intonational realization. Some of the features discussed here may be more prevalent in speech communities with an oral tradition and no established standard, but once spontaneous data become more common in research, the overall variability observed here is likely to prove comparable to that found in other speech communities. Thus the present corpus can be treated as an extreme example of variability which allows us to sharpen the intonational analysis toolkit. The lessons learned apply to the analysis of all languages, not exclusively to the present data, to Romani in particular, or to non-standardized languages like Romani.</p>
<p>I argue that the analysis presented here was easier to arrive at by not using a level of broad phonetic transcription as is often advocated (e.g., <xref ref-type="bibr" rid="B62">Hualde &amp; Prieto, 2016</xref>; Jun, 2005; <xref ref-type="bibr" rid="B68">Jun &amp; Fletcher, 2014</xref>). As discussed amply in Ladd (<xref ref-type="bibr" rid="B76">2008a, ch. 3</xref>), this approach represents one of the two main views about intonational analysis that played a part in the development of annotation systems, beginning with the original ToBI system for the prosodic annotation of English (<xref ref-type="bibr" rid="B106">Silverman et al., 1992</xref>). Specifically, the approach taken here is that advocated by some of the ToBI developers who took the position that the analysis of intonation using autosegmental-metrical representations is phonological in nature and therefore it need not faithfully represent every detail of the pitch contours (<xref ref-type="bibr" rid="B24">Beckman et al., 2005</xref>; <xref ref-type="bibr" rid="B76">Ladd, 2008a, p. 111</xref>; see also <xref ref-type="bibr" rid="B54">Gussenhoven, 2007</xref>, for similar views).</p>
<p>Adopting this position is not meant to denigrate the importance of phonetic detail. The value of phonetic detail in understanding speech has been noted for at least the past 30 years and is constantly affirmed by new evidence (see, <italic>inter alia</italic>, <xref ref-type="bibr" rid="B30">Browman &amp; Goldstein, 1992</xref>, on the repercussions of ignoring fine-grained phonetic detail in understanding allophonic variation; <xref ref-type="bibr" rid="B105">Scobbie et al., 2000</xref>, on covert contrast; <xref ref-type="bibr" rid="B42">Edwards et al., 2015</xref>, on the problems of ignoring phonetic detail in language acquisition). Intonation is no exception to this understanding. Almost a quarter of a century after the original ToBI system, it is undeniable that fine-grained phonetic detail is present in production and crucial for the processing of intonational categories (<xref ref-type="bibr" rid="B21">Barnes et al., 2012</xref>; <xref ref-type="bibr" rid="B33">Cangemi &amp; Grice, 2016</xref>; <xref ref-type="bibr" rid="B36">Cole &amp; Shattuck-Hufnagel, 2016</xref>; <xref ref-type="bibr" rid="B40">D&#8217;Imperio, 2000</xref>; <xref ref-type="bibr" rid="B41">D&#8217;Imperio et al., 2000</xref>; <xref ref-type="bibr" rid="B71">Knight, 2008</xref>; <xref ref-type="bibr" rid="B72">Knight &amp; Nolan, 2006</xref>). Thus the need for more research in this area is indisputable.</p>
<p>However, as advocated by Ladd (<xref ref-type="bibr" rid="B78">2011</xref>) for segmentals, investigating the details of phonetic realization neither necessitates recourse to broad phonetic transcriptions nor does it obviate the need for an abstract phonological analysis of intonation. Specifically, in his (<xref ref-type="bibr" rid="B78">2011</xref>) paper, Ladd argues in favour of such an abstract level of analysis and against a <italic>systematic phonetic level</italic>, the equivalent of a broad phonetic transcription. As Ladd shows, a systematic phonetic level is problematic as it converts one symbolic representation into another. Such more fine-grained symbolic representations may capture some details about realization but cannot capture all phonetic detail as shown by a large body of research of the past 30 years. To give but one example, timing is an important aspect of phonetic realization that no symbolic representation can capture, by definition (<xref ref-type="bibr" rid="B97">Port &amp; Leary, 2005</xref>).</p>
<p>The problem can be illustrated by first using a segmental example. At the phonological level, it is generally agreed that English has a phoneme /k/. At a systematic phonetic level, several allophones may be recognized, depending on a researcher&#8217;s emphasis on a particular aspect of realization; e.g., the detailed descriptions of Cruttenden (<xref ref-type="bibr" rid="B37">1994, pp. 138&#8211;157</xref>) and Ladefoged and Johnson (<xref ref-type="bibr" rid="B81">2011, pp. 57&#8211;65</xref>) focus in turn on aspiration, place of articulation, and type of release. Based on such descriptions, an aspirated (long-lag VOT) and an unaspirated (short-lag VOT) allophone are typically recognized for /k/, [k<sup>&#688;</sup>] and [k] respectively (cf. Section 3.2.). If emphasis is placed instead on place of articulation, [k], [&#107;&#799;], and [&#107;&#800;] allophones may be postulated (cf. <xref ref-type="bibr" rid="B37">Cruttenden, 1994, p. 153</xref>). Together VOT length and place of articulation would yield six /k/ allophones (possibly more if aspiration and place of articulation are combined with compatible types of release). However, these allophones (or any other for that matter) would not do justice to the attested variation in the realization of English /k/: VOT varies gradiently based on stress, quality of the following vowel, position in the foot, word, and phrase, and even on dialect (e.g., <xref ref-type="bibr" rid="B37">Cruttenden, 1994, pp. 140&#8211;142</xref>; <xref ref-type="bibr" rid="B69">Keating, 1984</xref>; <xref ref-type="bibr" rid="B109">Stuart-Smith et al., 2015</xref>); the exact place of articulation of /k/ is also different for each following vowel (<xref ref-type="bibr" rid="B37">Cruttenden, 1994, p. 153</xref>). This means that what is represented in a broad phonetic transcription &#8212; the realizations typically referred to as <italic>allophones</italic> &#8212; will be incomplete and arbitrary. As Browman and Goldstein (<xref ref-type="bibr" rid="B30">1992, p. 164</xref>) note: &#8220;many allophonic differences are just quantitative differences that are large enough that phoneticians/phonologists have been able to notice them, and to relate them to distinctive differences in other languages.&#8221;</p>
<p>By extension, a systematic, broad phonetic representation of intonation can only amount to an arbitrary collection of allotones without capturing the full gamut of variation. This is in fact explicitly noted by Hualde and Prieto (<xref ref-type="bibr" rid="B62">2016</xref>) who define broad phonetic transcription as &#8220;a form of transcription that includes <italic>a certain amount of redundant, phonologically non-contrastive detail</italic> that is nevertheless a systematic aspect of the language [emphasis added].&#8221; <italic>A certain amount</italic> is precisely the problem with such a system, as it is not clear how this amount can be determined (cf. <xref ref-type="bibr" rid="B33">Cangemi &amp; Grice, 2016</xref>, for similar arguments). For instance, Hualde and Prieto (<xref ref-type="bibr" rid="B62">2016</xref>: Figure <xref ref-type="fig" rid="F5">5</xref>) use !H% to phonetically transcribe an underlying L% boundary tone which, being undershot due to tonal crowding, is scaled higher than typical (by approximately 20 Hz). However, in that same figure, the H* of the L+H* pitch accent is also undershot, being scaled lower by approximately 20 Hz as well, but this change is not transcribed. Similarly arbitrary decisions could have been made for the Romani data presented here had a level of broad phonetic transcription been used. As Browman and Goldstein (<xref ref-type="bibr" rid="B30">1992</xref>) note, attention might have been paid to variants that have been used as distinctive tonal elements in other languages; !H% used by Hualde and Prieto to indicate an undershot L% is such an example.</p>
<p>It is this abritrariness of broad phonetic transcriptions that drives the position that abstract representations are more successfully combined with exemplars, detailed traces of phonetic realization (<xref ref-type="bibr" rid="B93">Pierrehumbert, 2002</xref>). As Beckman et al. (<xref ref-type="bibr" rid="B25">2007</xref>) have shown, both these levels &#8212; which can be loosely equated to phonological and phonetic &#8212; play a part in speech production and perception. What is doubtful, however, is that an intermediate systematic phonetic level plays a useful role either in linguistic behaviour or linguistic analysis (<xref ref-type="bibr" rid="B78">Ladd, 2011</xref>; <xref ref-type="bibr" rid="B94">Pierrehumbert et al., 2000</xref>; for similar arguments, see also <xref ref-type="bibr" rid="B36">Cole &amp; Shattuck-Hufnagel, 2016</xref>). If this applies to segmentals, then it is unclear why something different is advocated for intonation. Based on the above, it is clear that the issue is not whether a broad or a narrow transcription of intonation is to be preferred, while discussion cannot be fruitfully focused on the level of detail to be transcribed.<xref ref-type="fn" rid="n6">6</xref> It is not possible for any type of <italic>transcription</italic> to capture the full gamut of possible variability, while at the same time, using an intermediate systematic phonetic level can stop researchers from capturing essential generalizations (<xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>; <xref ref-type="bibr" rid="B30">Browman &amp; Goldstein, 1992</xref>).</p>
</sec>
<sec>
<title>5.3 The typology of intonation</title>
<p>The need for typological comparisons is an argument that has been put forward in favour of more surface faithful and detailed representations of intonation. As noted earlier, typological research is said to be hindered when similar phenomena are represented in different ways across languages (<xref ref-type="bibr" rid="B76">Ladd, 2008a, pp. 107&#8211;119</xref>, <xref ref-type="bibr" rid="B77">2008b</xref>; Prieto &amp; Hualde, 2016). At first glance, this seems like a legitimate concern. There are several elements of this argument, however, that warrant further scrutiny. First, it is not clear what kind of typology would require such consensus among representations. The typology envisaged either by Hyman (<xref ref-type="bibr" rid="B63">2006</xref>) or Beckman and Venditti (<xref ref-type="bibr" rid="B27">2011</xref>), to take two very different views, is not concerned with whether a system has a LH* accent and another a L+&lt;H*, but rather with the origin and function of tones. As Hyman (<xref ref-type="bibr" rid="B63">2006</xref>) notes, any <italic>phonological</italic> typology must deal <italic>not</italic> with surface phonetic details but rather with the analytical categories used to make sense of these details in a given linguistic system. Thus Romani would be classed as a language that uses only postlexical tones (intonation) in combination with stress. For a typology of this sort, more generic categories would work better to bring a cross-linguistic understanding about; but generic categories are unlikely to be phonetically transparent.</p>
<p>If, on the other hand, a phonetic typology is envisaged, then details are better captured in terms of algorithms or patterns of realization rather than by a detailed but still symbolic notation which, as shown in Section 5.2., is unlikely to adequately capture all variability in realization. Peak delay is a good example of the inadequacy of a symbolic system in capturing commonalities that would be of use in constructing a phonetic typology: if peak delay is a parameter to encode, how far from the onset of a stressed vowel should a peak be before an accent is annotated as having a delayed peak? In answering this question one needs to consider the fact that peak location is only the outcome of an algorithm and thus only an approximation to begin with (<xref ref-type="bibr" rid="B27">Beckman &amp; Venditti, 2011</xref>; <xref ref-type="bibr" rid="B73">Kochanski, 2010</xref>). Further, as shown in more detail below, the answer is clearly related to the system to which the accent belongs: if all peaks are systematically delayed, is it worth annotating delay at all? What if, like in Romani or Neapolitan Italian (<xref ref-type="bibr" rid="B33">Cangemi &amp; Grice, 2016</xref>), peaks show substantial variability in alignment? Similar arguments apply to the transcription of undershoot: how far from typical must a given tone&#8217;s scaling be before it is annotated? Can general criteria be established or should undershoot be defined for each speaker separately based on their pitch range and if so, how? If questions like these cannot be answered in a straightforward way &#8212; both because of logistical issues to do with how we measure turning points and define scaling relations, and because the answers to these questions cannot possibly be the same for all languages &#8212; we need to question the usefulness of such a level of transcription.</p>
<p>The issue of how to analyze linguistic systems and do typological comparisons is of concern to typologists in general. Some argue, like Ladd (<xref ref-type="bibr" rid="B76">2008a, pp. 107&#8211;119</xref>) or Hualde and Prieto (<xref ref-type="bibr" rid="B62">2016</xref>), that we need a predetermined set of categories into which to fit the elements of different systems. Others like Haspelmath (<xref ref-type="bibr" rid="B57">2010</xref>, <xref ref-type="bibr" rid="B58">2015</xref>) argue that a typology which relies on a limited set of categories from which all languages choose is unsatisfactory for many reasons. An obvious one is that such categories can be unnecessarily restrictive and may fail to capture essential generalizations (cf. <xref ref-type="bibr" rid="B58">Haspelmath, 2015, on clitics</xref>). This is particularly likely to be true in the field of intonational phonology, as only a fraction of languages have been adequately described and thus the whole gamut of possibilities in terms of the organization of prosodic features and their realization is simply unknown. The proposal by Hualde and Prieto (<xref ref-type="bibr" rid="B62">2016</xref>) illustrates this point. The authors provide a series of five labels for bitonal accents (H+L*, H*+L, L+H*, L+&lt;H*, L*+H) and propose canonical realizations for them. However, it is not certain that these five labels are sufficient to adequately capture all possible pitch accents researchers are likely to encounter as more languages are analyzed. Hualde and Prieto acknowledge that these labels should be broad enough to cover differences in realization, but this statement in itself implies that the essential categories are determined. This carries precisely the risk discussed by Haspelmath (<xref ref-type="bibr" rid="B57">2010</xref>, <xref ref-type="bibr" rid="B58">2015</xref>).</p>
<p>To avoid such problems, typologists argue that one can have recourse to <italic>comparative concepts</italic>, which can be used for cross-linguistic comparison, while recognizing that language-specific categories are needed to account for phenomena specific to each language (<xref ref-type="bibr" rid="B57">Haspelmath 2010</xref>, <xref ref-type="bibr" rid="B58">2015</xref>). What Ladd (<xref ref-type="bibr" rid="B76">2008a p. 110</xref>) calls &#8220;sustained level phrase-final pitch&#8221; could be such a concept when it comes to intonation. Pierrehumbert (<xref ref-type="bibr" rid="B92">1980</xref>) analyzed sustained level phrase-final pitch in English as a H-L% sequence of edge tones. In other analyses of English, however, it is argued to reflect the absence of a specific boundary tone (<xref ref-type="bibr" rid="B74">Ladd, 1983</xref>; <xref ref-type="bibr" rid="B48">Grabe, 1998, ch. 4</xref>, following <xref ref-type="bibr" rid="B51">Gussenhoven, 1984</xref>). In turn, the absence of a boundary tone is notated in some analyses as 0% (<xref ref-type="bibr" rid="B48">Grabe, 1998</xref>), or by not positing a tone at all, as in the German ToBI system, GToBI, in which sustained level phrase-final pitch is annotated as H-% (<xref ref-type="bibr" rid="B49">Grice et al., 2005</xref>). Arvaniti and Baltazani (<xref ref-type="bibr" rid="B9">2005</xref>), on the other hand, analyze sustained level phrase-final pitch in Greek as !H-!H%. Differences like these are seen by Ladd (<xref ref-type="bibr" rid="B76">2008a, pp. 107&#8211;119</xref>) as a problem. Ladd argues that sustained level phrase-final pitch is &#8220;on the face of it, a similar intonational phenomenon in different languages&#8221; (<xref ref-type="bibr" rid="B76">2008a, p. 110</xref>) and thus it should be presented in a similar way in all of them, because different representations can lead to the conclusion that languages differ more than they really do. Ladd&#8217;s point is well taken; his discussion, however, glosses over differences that relate to system-internal relationships between tonal elements in the languages he considers. Yet, the representation of sustained level phrase-final pitch (or any other intonational phenomenon, for that matter) does depend on the overall system of the language under analysis; by glossing over this critical point, Ladd makes the different representations appear utterly arbitrary, though they are motivated by system-internal consistency. This can be clearly seen if one compares English and Greek.</p>
<p>In Pierrehumbert&#8217;s (<xref ref-type="bibr" rid="B92">1980</xref>) analysis, sustained level phrase-final pitch comes about in the following manner. The H- of the H-L% configuration is downstepped because it is preceded by a H L sequence of tones (a H*+L accent to be exact); this is based on a more general tenet according to which all HLH tonal sequences trigger downstep of the second H tone (<xref ref-type="bibr" rid="B92">Pierrehumbert, 1980, p. 139</xref>). The downstepping of H- is explained somewhat differently in the revised analysis of Beckman and Pierrehumbert (<xref ref-type="bibr" rid="B26">1986</xref>), in which all bitonal accents are said to trigger downstep independently of the sequence of tones involved. Finally L% is upstepped because it follows a H- phrase accent; this solution is possible because in English H-L% sequences in which L% is fully scaled are not attested (but see <xref ref-type="bibr" rid="B55">Gussenhoven, 2016</xref>). Now in Greek, sustained level phrase-final pitch follows either a L*+H or L+H* pitch accent, depending on the melody. Crucially, there is no evidence that L*+H or L+H* triggers downstep in Greek (<xref ref-type="bibr" rid="B4">Arvaniti, 2003</xref>; <xref ref-type="bibr" rid="B9">Arvaniti &amp; Baltazani, 2005</xref>). In addition there is no HLH sequence on which downstep would apply.<xref ref-type="fn" rid="n7">7</xref> Thus, neither of the two explanations of downstep used for English is possible in Greek, nor is there any other context-related reason for the downstep. This leads to the conclusion that downstep has to be treated as an independent feature in Greek (as also argued for English in <xref ref-type="bibr" rid="B74">Ladd, 1983</xref>, and for Dutch in <xref ref-type="bibr" rid="B53">Gussenhoven, 2005</xref>). Further, unlike English, sequences of H-L% without L% upstep are attested in Greek (<xref ref-type="bibr" rid="B15">Arvaniti et al., 2006b</xref>), making a L%-upstep rule like that of English equally unsuitable for Greek. In short, sustained level phrase-final pitch in Greek cannot be analyzed in the same way as sustained level phrase-final pitch in English. One can of course question whether H-L% is the only possible or optimal way of analysing sustained level phrase-final pitch in English; e.g., Gussenhoven (<xref ref-type="bibr" rid="B55">2016</xref>) presents cogent arguments against this analysis. This, however, remains an analytical decision <italic>about English</italic> and as such it should have little bearing on how sustained level phrase-final pitch is analyzed in Greek or any other language. As this example demonstrates, different decisions are the outcome of different system requirements. Among the differences appears to be the fact that the tonal space is carved up in ways that make it impossible to use just L and H tones for the analysis of all linguistic systems. Indeed any explicit use of the downstep feature argues in essence for a system with three levels (<xref ref-type="bibr" rid="B31">Brugos et al., 2006</xref>; <xref ref-type="bibr" rid="B74">Ladd, 1983</xref>; cf. <xref ref-type="bibr" rid="B82">Liberman, 1975</xref>).</p>
<p>At best then, one could argue that Arvaniti and Baltazani (<xref ref-type="bibr" rid="B9">2005</xref>) could have followed the convention established by GToBI and used !H-% instead of !H-!H% to indicate the lack of change in pitch (cf. <xref ref-type="bibr" rid="B49">Grice et al., 2005</xref>).<xref ref-type="fn" rid="n8">8</xref> This however, is a simple question of notation, not a question of analysis or typology. On the other hand, and this is a crucial difference, an analysis whereby sustained level phrase-final pitch is represented either by a 0% boundary tone (as in <xref ref-type="bibr" rid="B48">Grabe, 1998</xref>) or no boundary tone at all (as in <xref ref-type="bibr" rid="B52">Gussenhoven, 2004, pp. 313&#8211;315</xref>) would require altogether different analytical decisions. As indicated in Section 4.2.3., such an analysis would require that the pitch accent of the melody includes the downstep which in the GRToBI analysis is represented as a sequence of two distinct tonal events, !H- and !H%, both independent of the pitch accent. Whether one or the other theoretical position is superior is beyond the scope of this paper, though it is likely that each is better suited for some languages than others.</p>
<p>The discussion above should serve to highlight the fact that differences among analyses are not all qualitatively the same. The distinctions between them should be acknowledged, as some are genuine problems with straightforward solutions and others are part of the nature of research itself. The differences are of three types which are discussed below primarily in relation to AM analyses of the vocative chant in a variety of languages (see Table <xref ref-type="table" rid="T2">2</xref>). The vocative chant is used here because it is the most characteristic use of sustained level phrase-final pitch which, as noted above, has been a matter of some debate.</p>
<list list-type="roman-lower">
<list-item>
<p>Differences between intonational systems. Differences between systems arise for two reasons: first, melodies that are similar in form and function may still show differences substantial enough to warrant distinct representations; second, distinct representations may be required for the sake of analytical consistency. The former type is illustrated by Frota (<xref ref-type="bibr" rid="B45">2016</xref>), regarding the rise-fall contour associated with narrow focus in both Portuguese and Catalan. Frota shows that, despite superficial similarities, both production and perception data indicate that the former is an off-ramp H*+L and the latter an on-ramp L+H* pitch accent. On the other hand, the representation of sustained level phrase-final pitch in Greek discussed above exemplifies the system-internal considerations that force a particular analysis.</p>
</list-item>
<list-item>
<p>Distinct analytical positions. As noted in Section 4.2.3. and above, decisions about how to carve up a melody into distinct tonal events have consequences for their representation. This is the reason why the Dutch vocative chant is analyzed without recourse to an edge tone in Gussenhoven (<xref ref-type="bibr" rid="B53">2005</xref>): in his analyses, the drop from a high to mid-level pitch (which is then sustained) is analyzed as part of the H*!H pitch accent. H*!H % may indeed be best for Dutch as it reflects the fact that the melody applies to successive feet, when available, a behaviour typical of pitch accents (<xref ref-type="bibr" rid="B53">Gussenhoven, 2005</xref>; see <xref ref-type="bibr" rid="B50">Grice et al., 2000</xref>, for an alternative analysis). This type of analysis is not suitable for Greek or Polish, however, since both languages have only one level of stress, a metrical difference that makes iteration of the melody impossible (<xref ref-type="bibr" rid="B5">Arvaniti, 2007a</xref>; <xref ref-type="bibr" rid="B9">Arvaniti &amp; Baltazani, 2005</xref>; <xref ref-type="bibr" rid="B16">Arvaniti et al., 2016</xref>).</p>
</list-item>
<list-item>
<p>Notational differences. Notational differences are evident in the representations of the vocative chant in Table <xref ref-type="table" rid="T2">2</xref>; e.g. L+H* and LH* represent pitch accents with very similar characteristics; !H-0% and H-% arguably represent the same thing, sustained mid-level pitch as a reflex of phrasal tones.</p>
</list-item>
</list>
<table-wrap id="T2">
<label>Table 2</label>
<caption>
<p>AM representations of sustained phase-final pitch as used in the vocative chant.</p>
</caption>
<table>
<tr>
<th align="left">Language</th>
<th align="left">Pitch accent</th>
<th align="left">Phrasal tones</th>
</tr>
<tr>
<td colspan="3">
<hr/></td>
</tr>
<tr>
<td align="left">Catalan (<xref ref-type="bibr" rid="B28">Borr&#224;s-Comes et al., 2015</xref>)</td>
<td align="left">L+H*</td>
<td align="left">!H%</td>
</tr>
<tr>
<td align="left">Dutch (<xref ref-type="bibr" rid="B53">Gussenhoven, 2005</xref>)</td>
<td align="left">H*!H*</td>
<td align="left">%</td>
</tr>
<tr>
<td align="left">English (<xref ref-type="bibr" rid="B31">Brugos et al., 2006</xref>)</td>
<td align="left">H*</td>
<td align="left">!H-L%</td>
</tr>
<tr>
<td align="left">German (<xref ref-type="bibr" rid="B49">Grice et al., 2005</xref>)</td>
<td align="left">L+H*</td>
<td align="left">H-%</td>
</tr>
<tr>
<td align="left">Greek (<xref ref-type="bibr" rid="B9">Arvaniti &amp; Baltazani, 2005</xref>)</td>
<td align="left">L*+H</td>
<td align="left">!H-!H%</td>
</tr>
<tr>
<td align="left">Hungarian (<xref ref-type="bibr" rid="B114">Varga, 2008</xref>)</td>
<td align="left">H*</td>
<td align="left">!H-0%</td>
</tr>
<tr>
<td align="left">Polish (<xref ref-type="bibr" rid="B16">Arvaniti et al., 2016</xref>)</td>
<td align="left">LH*</td>
<td align="left">!H-%</td>
</tr>
<tr>
<td align="left">Portuguese (<xref ref-type="bibr" rid="B46">Frota et al., 2015</xref>)</td>
<td align="left">(L+)H*</td>
<td align="left">!H%</td>
</tr>
</table>
</table-wrap>
<p>The three types of differences discussed above cannot be approached in the same way. Differences between systems should be accepted as inevitable. Languages cannot be expected to have the same tonal inventory, use the same melodies, carve up the tonal space in the same manner, exhibit the same interactions between tones, or otherwise realize the same phonological entities in the same manner in all contexts (cf. the differences between Portuguese and Catalan reported in <xref ref-type="bibr" rid="B45">Frota, 2016</xref>). The prosodic type of the language in question and the interaction between metrical and tonal structure are additional sources of cross-linguistic variation. Such differences, as argued above, may lead by necessity to very different analyses, if those analyses are to be internally consistent.</p>
<p>On the other hand, disagreements in notation can be resolved relatively easily by agreeing on a set of consistent conventions. Such agreement could be reached on how to annotate a sequence of two identical edge tones: L-L%, L-% or L-0% etc. (but see Sections 5.4 and 5.5. below). Similarly, agreement should be possible on whether multi-tonal accents are best represented with the plus sign between tones or not (e.g., L+H* or LH*), or whether Jun and Fletcher&#8217;s (<xref ref-type="bibr" rid="B68">2014</xref>) proposal to distinguish the two in a principled manner is to be preferred. It is important to keep in mind, however, that such differences are trivial (however intimidating they may be to non-initiates).</p>
<p>Distinguishing between notational and analytical disagreements is crucial for any attempt to standardize AM representations, especially as it appears that the two types of disagreement are sometimes overlooked: Hualde &amp; Prieto (<xref ref-type="bibr" rid="B62">2016</xref>) treat the difference between !H% in the analysis of the Portuguese vocative chant (<xref ref-type="bibr" rid="B46">Frota et al., 2015</xref>) and !H-% in the German equivalent (<xref ref-type="bibr" rid="B49">Grice et al., 2005</xref>) as being on a par with the difference between the German !H-% and the !H-!H% used in Greek (<xref ref-type="bibr" rid="B9">Arvaniti &amp; Baltazani, 2005</xref>). However, the difference between the Greek and German analyses is one of convention (a notational difference) while that between Portuguese and German reflects different analytical decisions about edge tones: Frota et al.&#8217;s analysis of Portuguese relies on boundary tones, while Grice et al.&#8217;s analysis of German posits both phrase accents and boundary tones. Similarly, the difference between Ladd and Schepman&#8217;s (<xref ref-type="bibr" rid="B80">2003</xref>) analysis of English rising accents as (L+H)* and the use of H* for Romani is not a difference in notation; rather, it is a different analytical approach to the role and significance of the initial rise in such accents.</p>
<p>Analytical differences are not easy to resolve as they reflect different approaches to phenomena, often coupled with different requirements of the systems under analysis. Nevertheless, agreement in analytical decisions appears to be a desideratum for some; e.g., Hualde &amp; Prieto (<xref ref-type="bibr" rid="B62">2016</xref>) talk of &#8220;the potential use of a generally accepted set of intonational labels and phonetic implementation rules that can be common across languages&#8221;. Such a goal, however, would not only force all languages onto a phonetic Procrustean bed, but would also require that all researchers espouse the exact same principles and solutions to problems of analysis. Such homogeneity of opinion would be detrimental to scientific inquiry, and very unlikely to be achieved.</p>
<p>Though analytical differences among researchers will and should persist, useful progress could be made by working towards a generally agreed set of criteria and diagnostic tests that would allow researchers to evaluate alternative analyses for the same linguistic system on a consistent basis. Examples of recent research along these lines include Peters et al. (<xref ref-type="bibr" rid="B91">2015</xref>), Ritter and Grice (<xref ref-type="bibr" rid="B103">2015</xref>), and Gussenhoven (<xref ref-type="bibr" rid="B55">2016</xref>). As these studies indicate, criteria could relate, on the one hand, to levels of adequacy that analyses must meet, and on the other, to the empirical evidence that must support an analysis. Such criteria could include types of empirical evidence required to determine whether one or two types of edge tones are needed for the analysis of a given language, whether to posit one or more types of rising accents and what their tonal composition might be. Focusing on the development of such diagnostic tests, on the one hand, and on standardizing notation where appropriate, on the other, should help resolve many points of disagreement among analyses.</p>
</sec>
<sec>
<title>5.4 Phonetic transparency in intonation</title>
<p>Another reason put forward for more similarity in cross-linguistic representations of intonation is the need for phonetic transparency (<xref ref-type="bibr" rid="B76">Ladd, 2008a, p. 112</xref>). As Ladd concedes, however, there are some problems with this argument, in that phonetic transparency can complicate rather than facilitate comparisons across linguistic varieties. Ladd uses the vowel system of Scottish English to illustrate this difficulty: Scottish English does not have a contrast between /&#650;/ and /u/, and the vowel used in place of both is best transcribed as [&#649;]. Thus, neither /&#650;/ nor /u/ used in standard analyses of English is a good representation for the high mid central Scottish vowel; however, if, in the name of phonetic transparency, both /&#650;/ and /u/were to be replaced by /&#649;/ in the analysis of Scottish English, it would be difficult to compare the Scottish English vowel system with that of Southern Standard British English.<xref ref-type="fn" rid="n9">9</xref> Ladd notes, however, that, while neither /&#650;/ nor /u/ is an ideal representation for the high central Scottish English vowel, no one would consider representing the vowel of <italic>brick</italic> or <italic>break</italic> using /&#650;/. In other words, there is some largely agreed upon phonetic substance related to these symbolic representations.</p>
<p>To my knowledge at least, the same applies to analyses of intonation. There are no analyses in which high pitch is represented by a L tone and low pitch by a H tone, a counterintuitive analytical decision equivalent to Ladd&#8217;s <italic>brick</italic> transcribed with /&#650;/. The main point of disagreement across intonational analyses concerns sustained level phrase-final pitch (essentially mid-level pitch). This is due partly to differences among languages, as noted in Section 5.3., and partly to historical reasons which resulted in the adoption to a two-tone system forcing some rather cumbersome representations of pitch that is neither low nor high but is contrastive (<xref ref-type="bibr" rid="B7">Arvaniti, 2011</xref>). It is no coincidence that this is a main topic of scrutiny for four out of six papers in this collection (Arvaniti, 2016; <xref ref-type="bibr" rid="B45">Frota, 2016</xref>; <xref ref-type="bibr" rid="B55">Gussenhoven, 2016</xref>; Prieto and Hualde, 2016).</p>
<p>The fact that phonological representations of intonation are not phonetically arbitrary is illustrated in Table <xref ref-type="table" rid="T2">2</xref> which lists the representations of the vocative chant, a melody that according to Ladd shows &#8220;striking similarity across the languages of Europe&#8221; (<xref ref-type="bibr" rid="B76">Ladd, 2008a, p. 119</xref>). As can be seen in Table <xref ref-type="table" rid="T2">2</xref>, all analyses involve a rising or high accent followed by sustained mid-level pitch. Differences in representation may seem overwhelming at first glance, but do not really obscure similarities and are not any more arbitrary than any phonological analysis of segments. Granted, one has to know that in the English system !H-L% involves an upstep of L%, but this is no different from having to learn that /p/ in English is aspirated in most contexts or that /b/ is rarely fully voiced (<xref ref-type="bibr" rid="B69">Keating, 1984</xref>) and thus that the symbols /p/ and /b/ do not represent quite the same sounds in English and French. What is noteworthy is that these types of discrepancies have long been accepted in segmental phonology, but are still treated as highly undesirable in the analysis of intonation. As I have argued elsewhere, one possible explanation is that intonation is not seen as being on a par with the rest of phonological structure even by those who study it (<xref ref-type="bibr" rid="B6">Arvaniti, 2007b</xref>). This tendency is probably reinforced by the relative phonetic transparency of L and H which forces a phonetic interpretation of phonological representations of intonation, impossible for abstract symbols like /p/ or /b/.</p>
</sec>
<sec>
<title>5.5 Dialectology, cross-linguistic comparisons, and the choice of categories</title>
<p>Dialectological research is another argument that has been used in support of a broad phonetic level of intonation transcription (<xref ref-type="bibr" rid="B62">Hualde &amp; Prieto, 2016</xref>). Yet such transcriptions are now largely abandoned by dialectologists for the reasons discussed by Ladd (<xref ref-type="bibr" rid="B76">2008a, pp. 110&#8211;115</xref>) and briefly in Section 5.3. Following Wells (<xref ref-type="bibr" rid="B116">1982</xref>), instead of talking about /&#650;/ or /&#594;/, sociolinguists working on English talk about the FOOT and the LOT vowel respectively, a practice indicating a level of abstraction similar to that advocated here for intonation; talking about the FOOT vowel or the LOT vowel obviates the need to label the phonetic substance of these vowels but does allow for fruitful comparisons.</p>
<p>One thing to notice about this practice, however, is the cultural hegemony it reflects. The list of words used is based on categories that come from the system of Standard Southern British English. As luck would have it, it is the English vowel system with the largest number of vowel contrasts and thus it serves English dialectology well, but one wonders what that list of words would have been had it first been proposed by a speaker from Los Angeles or Newcastle; in the former case, there would be no separate entries for THOUGHT and LOT, while in the latter STRUT would be missing instead. Such biases, which are inevitable, add another layer of arbitrariness. Problems of this sort are inevitably compounded when a system of broad phonetic transcription is used precisely because phonetic substance cannot be left unspecified.</p>
<p>Indeed, problems do arise in dialectological comparisons when researchers opt for overly transparent phonological presentations. One such case is the proposal of Ladd and Schepman (<xref ref-type="bibr" rid="B80">2003</xref>) to collapse H* and L+H* into (L+H)*. As noted in Section 3.2., Arvaniti and Garding (<xref ref-type="bibr" rid="B11">2007</xref>) have shown that there is dialectal variation in the realization of these accents: in their study, speakers from California made a consistent distinction between H* and L+H*, using H* for new information and L+H* for contrastive focus, as suggested by Pierrehumbert (<xref ref-type="bibr" rid="B92">1980</xref>) and Pierrehumbert &amp; Hirschberg (<xref ref-type="bibr" rid="B95">1990</xref>). Speakers from Minnesota, on the other hand, clearly had one pitch accent, L+H*, and relied on scaling to indicate the difference between new and contrastive information, as Ladd and Schepman (<xref ref-type="bibr" rid="B80">2003</xref>) would predict. Given these dialectal differences, collapsing the two categories in all descriptions of English intonation would be counterproductive as it would obscure a more important difference across dialects of English: the presence (or absence) of the L+H* vs. H* contrast. Doing so would be equivalent to using /&#649;/ for the analysis of all English varieties because Scottish has this vowel, or using only /&#596;/ for CAUGHT and LOT because Western varieties of US English have merged these vowels into one.</p>
<p>There are some additional concerns with respect to dialectology that go beyond cultural hegemony. A common transcription system implies that varieties of a language share some common core. In languages with well accepted and known standardized forms, this may be desirable and realistic and may have some psychological reality as well in that non-standard speakers are likely to be familiar with the standard. Experience, however, suggests that this does not apply to all speech communities, even those with highly codified standards: thus, British speakers are far more aware of a UK-wide English standard and are familiar with terms such as <italic>RP, Queen&#8217;s English</italic>, and <italic>BBC English</italic>; for U.S. speakers, on the other hand, concepts like <italic>Mainstream American English</italic> or <italic>General American English</italic> hold little reality. This makes the enterprise of a common system possibly useful for linguists but of little psychological validity. This is all the more so for speakers like the Roma in the present study who are not familiar with a standard form of their language. In such circumstances, it would be highly unrealistic to posit that a common system for all Romani varieties must be used either for segmental or prosodic analysis, as such a system has no bearing on specific varieties and speakers. This state of affairs is likely to hold for speakers of other languages without a standard and without a written and schooling tradition. This in turn means that while we discuss phonetic transparency, we make analytical decisions that fit one variety better than others, as would happen if the present analysis were to be made the base of intonation analysis in other Romani varieties.</p>
<p>The tendency for some linguistic varieties to take priority is implicit in cross-linguistic work as well; e.g., Hualde et al. (<xref ref-type="bibr" rid="B61">2002</xref>) argue that Lekeitio Basque is like Japanese; if Basque had been analyzed first it would be Japanese that would have to fit the Basque type. Although the similarities in these particular systems may render this difference trivial, issues of precedence can have consequences for other analyses: it is undeniable that many analytical decisions in intonation have been the way they are because of the influence of English.</p>
</sec>
</sec>
<sec>
<title>6 Conclusion</title>
<p>In conclusion, the corpus presented here shows variability on a scale rarely encountered in data from educated monolingual speakers of standardized languages, though presumably common in many non-standardized linguistic varieties, particularly those showing extensive contact. Variability on this scale poses challenges for intonational analysis and highlights the importance of distinguishing between phonetic realization and phonological representation during analysis and determining intonational phonology on the basis of <italic>meaningful</italic> contrasts as in the rest of a language&#8217;s phonological system. Though the need to adhere to these principles may be more obvious under conditions of variability like those discussed here, the argument made is that the principles would be useful in the intonational analysis of all languages. Such an analysis can be usefully and fruitfully compared with analyses of other languages leading to successful typological comparisons. This can be achieved without recourse to an intermediate level of broad phonetic transcription which cannot do justice to the full gamut of variability in any data, but will inevitably focus on some variable elements whether they are especially significant or not. Doing so may well obscure real cross-linguistic similarities and lead to researchers missing both significant generalizations and the opportunity to explore the full gamut of prosodic variation present in the world&#8217;s languages.</p>
</sec>
</body>
<back>
<fn-group>
<fn id="n1">
<p>This statement is not meant to imply that there are no exceptions, such as participants who cannot read a corpus of isolated sentences in a consistent style or speaking rate, or using a constant melody. But typically such speakers are the exception and not included in analysis.</p>
</fn>
<fn id="n2">
<p>According to FRA and UNDP, in 2011 less than 10% of Roma children in Greece attended pre-school or kindergarten, while just over 35% of Roma children aged 7&#8211;15 attended school; less than 1% of the Roma population overall had completed upper secondary education.</p>
</fn>
<fn id="n3">
<p>Though the majority of the data comes from spontaneous speech, many illustrations here are from QUIS, as the data were less noisy and produced with clearer speech.</p>
</fn>
<fn id="n4">
<p>A similar result can be achieved using Rapid Prosody Transcription, a system in which boundaries are marked by lay participants; the results can be used to examine the acoustic parameters associated with high boundary scores (<xref ref-type="bibr" rid="B36">Cole &amp; Shattuck-Hufnagel, 2016</xref>).</p>
</fn>
<fn id="n5">
<p>It is possible that secondary association is needed to account for the realization of the LH% boundary tone attested with <italic>wh</italic>-questions. The corpus, however, contains too few questions with this pattern to allow for further analysis.</p>
</fn>
<fn id="n6">
<p>It is sometimes argued that broad phonetic transcriptions would be useful for applications, such as speech synthesis. It is not possible to address this point here in detail, but typically speech synthesis systems rely on mark-up languages and implementation rules that incorporate quantitative measures rather than symbolic representations alone (cf. <xref ref-type="bibr" rid="B60">Hirschberg, 2006</xref>; <xref ref-type="bibr" rid="B108">Sproat et al., 1998</xref>; <xref ref-type="bibr" rid="B113">van Santen et al., 1997</xref>; <xref ref-type="bibr" rid="B115">Venditti et al., 2008</xref>).</p>
</fn>
<fn id="n7">
<p>I do not consider here the fact that alternative representations of Greek accents could create a HLH sequence; rather, I assume that the representations of the accents are correct.</p>
</fn>
<fn id="n8">
<p>Even this representation may not be optimal for Greek, as !H% is needed independently of sustained level phrase-final pitch to represent the final tone in Greek <italic>wh</italic>-questions and related melodies (<xref ref-type="bibr" rid="B12">Arvaniti &amp; Ladd, 2009</xref>; <xref ref-type="bibr" rid="B10">Arvaniti et al., 2014</xref>; <xref ref-type="bibr" rid="B18">Baltazani, 2006</xref>).</p>
</fn>
<fn id="n9">
<p>Ladd does not address distributional, weight-related, and metrical criteria that could help determine the optimal (or least objectionable) phonological representation of Scottish [&#649;].</p>
</fn>
</fn-group>
<sec>
<title>Supplementary Files</title>
<p>For accompanying TextGrid, Pitch, and wav files, go to <inline-supplementary-material xlink:href="http://dx.doi.org/10.5334/labphon.14.smo">http://dx.doi.org/10.5334/labphon.14.smo</inline-supplementary-material>.</p>
</sec>
<ack>
<title>Acknowledgements</title>
<p>Many thanks to my collaborator on this project, Evangelia Adamou, who introduced me to Romani, and to the Roma consultants without whose generous participation this work would not have been possible. The recordings, glosses, and translations of the examples cited in this paper are by Evangelia Adamou. Thanks are also due to the editors and two anonymous reviewers for their detailed and insightful comments. The financial support of CNRS through the F&#233;d&#233;ration de Typologie et Universaux Linguistiques, and of an Illinois-WUN International Development Grant is hereby gratefully acknowledged.</p>
</ack>
<sec>
<title>Competing Interests</title>
<p>The author declares that they have no competing interests.</p>
</sec>
<ref-list>
<ref id="B1">
<label>1</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Adamou</surname>
<given-names>E.</given-names>
</name>
</person-group>
<article-title>Bilingual speech and language ecology in Greek Thrace: Romani and Pomak in contact with Turkish</article-title>
<source>Language in Society</source>
<year iso-8601-date="2010">2010</year>
<volume>39</volume>
<fpage>147</fpage>
<lpage>171</lpage>
<pub-id pub-id-type="doi">10.1017/S0047404510000035</pub-id>
</element-citation>
</ref>
<ref id="B2">
<label>2</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Adamou</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Illustrations of the IPA: Greek Thrace Xoraxane Romane</article-title>
<source>Journal of the International Phonetic Association</source>
<year iso-8601-date="2014">2014</year>
<volume>44</volume>
<issue>2</issue>
<fpage>223</fpage>
<lpage>231</lpage>
<pub-id pub-id-type="doi">10.1017/S0025100313000376</pub-id>
</element-citation>
</ref>
<ref id="B3">
<label>3</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Anderson</surname>
<given-names>A. H.</given-names>
</name>
<name>
<surname>Bader</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Bard</surname>
<given-names>E. G.</given-names>
</name>
<name>
<surname>Boyle</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Doherty</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Garrod</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Isard</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Kowtko</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>McAllister</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Miller</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Sotillo</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Thompson</surname>
<given-names>H. S.</given-names>
</name>
<name>
<surname>Weinert</surname>
<given-names>R.</given-names>
</name>
</person-group>
<article-title>The HCRC Map Task Corpus</article-title>
<source>Language and Speech</source>
<year iso-8601-date="1991">1991</year>
<volume>34</volume>
<fpage>351</fpage>
<lpage>366</lpage>
</element-citation>
</ref>
<ref id="B4">
<label>4</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Peak scaling in Greek and the role of declination</article-title>
<conf-name>Proceedings of XVth ICPhS</conf-name>
<conf-date>4&#8211;9 August 2003</conf-date>
<year iso-8601-date="2003">2003</year>
<conf-loc>Barcelona</conf-loc>
<fpage>2269</fpage>
<lpage>2272</lpage>
</element-citation>
</ref>
<ref id="B5">
<label>5</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Greek phonetics: The state of the art</article-title>
<source>Journal of Greek Linguistics</source>
<year iso-8601-date="2007a">2007a</year>
<volume>8</volume>
<fpage>97</fpage>
<lpage>208</lpage>
<pub-id pub-id-type="doi">10.1075/jgl.8.08arv</pub-id>
</element-citation>
</ref>
<ref id="B6">
<label>6</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>On the relationship between phonology and phonetics (or why phonetics is not phonology). Special Session: Between Meaning and Speech: On the Role of Communicative Functions, Representations and Articulations</article-title>
<conf-name>Proceedings of ICPhS XVI</conf-name>
<year iso-8601-date="2007b">2007b</year>
<fpage>19</fpage>
<lpage>24</lpage>
</element-citation>
</ref>
<ref id="B7">
<label>7</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>van Oostendorp</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ewen</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Hume</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Rice</surname>
<given-names>K.</given-names>
</name>
</person-group>
<chapter-title>The representation of intonation</chapter-title>
<source>Companion to Phonology</source>
<year iso-8601-date="2011">2011</year>
<publisher-loc>Malden, MA</publisher-loc>
<publisher-name>Wiley-Blackwell</publisher-name>
<fpage>757</fpage>
<lpage>780</lpage>
</element-citation>
</ref>
<ref id="B8">
<label>8</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Adamou</surname>
<given-names>E.</given-names>
</name>
</person-group>
<article-title>Focus expression in Romani</article-title>
<conf-name>28th West Coast Conference on Formal Linguistics</conf-name>
<year iso-8601-date="2011">2011</year>
<conf-sponsor>Cascadilla Proceedings Project</conf-sponsor>
</element-citation>
</ref>
<ref id="B9">
<label>9</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Baltazani</surname>
<given-names>M.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>Intonational analysis and prosodic annotation of Greek spoken corpora</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005">2005</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>84</fpage>
<lpage>117</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0004</pub-id>
</element-citation>
</ref>
<ref id="B10">
<label>10</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Baltazani</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Gryllia</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>The pragmatic interpretation of intonation in Greek wh-questions</article-title>
<conf-name>Proceedings of Speech Prosody 7</conf-name>
<conf-date>May 20&#8211;23, 2011</conf-date>
<year iso-8601-date="2014">2014</year>
<conf-loc>Dublin</conf-loc>
<comment>Retrieved from <uri>http://www.speechprosody2014.org</uri></comment>
</element-citation>
</ref>
<ref id="B11">
<label>11</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Garding</surname>
<given-names>G.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Cole</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Hualde</surname>
<given-names>J. I.</given-names>
</name>
</person-group>
<chapter-title>Dialectal variation in the rising accents of American English</chapter-title>
<source>Papers in Laboratory Phonology</source>
<year iso-8601-date="2007">2007</year>
<publisher-loc>Berlin, New York</publisher-loc>
<publisher-name>Mouton de Gruyter</publisher-name>
<fpage>547</fpage>
<lpage>576</lpage>
</element-citation>
</ref>
<ref id="B12">
<label>12</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<article-title>Greek wh-questions and the phonology of intonation</article-title>
<source>Phonology</source>
<year iso-8601-date="2009">2009</year>
<volume>26</volume>
<fpage>43</fpage>
<lpage>74</lpage>
<pub-id pub-id-type="doi">10.1017/S0952675709001717</pub-id>
</element-citation>
</ref>
<ref id="B13">
<label>13</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Mennen</surname>
<given-names>I.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Broe</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J.</given-names>
</name>
</person-group>
<chapter-title>What is a starred tone? Evidence from Greek</chapter-title>
<source>Papers in Laboratory Phonology V: Acquisition and the Lexicon</source>
<year iso-8601-date="2000">2000</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<fpage>119</fpage>
<lpage>131</lpage>
</element-citation>
</ref>
<ref id="B14">
<label>14</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Mennen</surname>
<given-names>I.</given-names>
</name>
</person-group>
<article-title>Tonal association and tonal alignment: evidence from Greek polar questions and contrastive statements</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2006a">2006a</year>
<volume>49</volume>
<fpage>421</fpage>
<lpage>450</lpage>
<pub-id pub-id-type="doi">10.1177/00238309060490040101</pub-id>
</element-citation>
</ref>
<ref id="B15">
<label>15</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Mennen</surname>
<given-names>I.</given-names>
</name>
</person-group>
<article-title>Phonetic effects of focus and &#8220;tonal crowding&#8221; in intonation: Evidence from Greek polar questions</article-title>
<source>Speech Communication</source>
<year iso-8601-date="2006b">2006b</year>
<volume>48</volume>
<fpage>667</fpage>
<lpage>696</lpage>
<pub-id pub-id-type="doi">10.1016/j.specom.2005.09.012</pub-id>
</element-citation>
</ref>
<ref id="B16">
<label>16</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Zygis</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Jaskula</surname>
<given-names>M.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>&#379;ygis</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Malisz</surname>
<given-names>Z.</given-names>
</name>
</person-group>
<article-title>The phonetics and phonology of the Polish calling contours</article-title>
<source>Phonetica</source>
<year iso-8601-date="2016">2016</year>
<volume>73</volume>
<issue>Special Issue</issue>
<comment>&#8220;Slavic perspectives on prosody&#8221;</comment>
</element-citation>
</ref>
<ref id="B17">
<label>17</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Bakker</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Yaron</surname>
<given-names>M.</given-names>
</name>
</person-group>
<source>Bibliography of modern Romani linguistics</source>
<year iso-8601-date="2003">2003</year>
<publisher-loc>Amsterdam</publisher-loc>
<publisher-name>Benjamins</publisher-name>
<pub-id pub-id-type="doi">10.1075/lisl.28</pub-id>
</element-citation>
</ref>
<ref id="B18">
<label>18</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Baltazani</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Intonation and pragmatic interpretation of negation in Greek</article-title>
<source>Journal of Pragmatics</source>
<year iso-8601-date="2006">2006</year>
<volume>38</volume>
<issue>10</issue>
<fpage>1658</fpage>
<lpage>1676</lpage>
<pub-id pub-id-type="doi">10.1016/j.pragma.2005.03.016</pub-id>
</element-citation>
</ref>
<ref id="B19">
<label>19</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Baltazani</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<article-title>Topic and focus intonation in Greek</article-title>
<source>Proceedings of the XIVth International Congress of Phonetic Sciences</source>
<year iso-8601-date="1999">1999</year>
<volume>2</volume>
<fpage>1305</fpage>
<lpage>1308</lpage>
</element-citation>
</ref>
<ref id="B20">
<label>20</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Barnes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Brugos</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Veilleux</surname>
<given-names>N.</given-names>
</name>
</person-group>
<article-title>The domain of realization of the L- phrase tone in American English</article-title>
<source>Proceedings of Speech Prosody</source>
<conf-date>2006</conf-date>
<year iso-8601-date="2006">2006</year>
<conf-sponsor>Dresden</conf-sponsor>
</element-citation>
</ref>
<ref id="B21">
<label>21</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Barnes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Veilleux</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Brugos</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Tonal center of gravity: A global approach to tonal implementation in a level-based intonational phonology</article-title>
<source>Laboratory Phonology</source>
<year iso-8601-date="2012">2012</year>
<volume>3</volume>
<issue>2</issue>
<fpage>337</fpage>
<lpage>383</lpage>
<pub-id pub-id-type="doi">10.1515/lp-2012-0017</pub-id>
</element-citation>
</ref>
<ref id="B22">
<label>22</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
</person-group>
<source>Stress and non-stress Accent</source>
<year iso-8601-date="1986">1986</year>
<publisher-loc>Dordrecht</publisher-loc>
<publisher-name>Foris</publisher-name>
<pub-id pub-id-type="doi">10.1515/9783110874020</pub-id>
</element-citation>
</ref>
<ref id="B23">
<label>23</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Edwards</surname>
<given-names>J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Keating</surname>
<given-names>P. A.</given-names>
</name>
</person-group>
<chapter-title>Articulatory evidence for differentiating stress categories</chapter-title>
<source>Phonological structure and phonetic form: Papers in Laboratory Phonology III</source>
<year iso-8601-date="1994">1994</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<fpage>7</fpage>
<lpage>33</lpage>
</element-citation>
</ref>
<ref id="B24">
<label>24</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>The original ToBI system and the evolution of the ToBI framework</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005">2005</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>9</fpage>
<lpage>54</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0002</pub-id>
</element-citation>
</ref>
<ref id="B25">
<label>25</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Munson</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Edwards</surname>
<given-names>J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Cole</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Hualde</surname>
<given-names>J. I.</given-names>
</name>
</person-group>
<chapter-title>The influence of vocabulary growth on developmental changes in types of phonological knowledge</chapter-title>
<source>Laboratory Phonology 9</source>
<year iso-8601-date="2007">2007</year>
<publisher-loc>Berlin, New York</publisher-loc>
<publisher-name>Mouton de Gruyter</publisher-name>
<fpage>241</fpage>
<lpage>264</lpage>
</element-citation>
</ref>
<ref id="B26">
<label>26</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>Intonational structure in Japanese and English</article-title>
<source>Phonology</source>
<year iso-8601-date="1986">1986</year>
<volume>3</volume>
<fpage>15</fpage>
<lpage>70</lpage>
</element-citation>
</ref>
<ref id="B27">
<label>27</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Venditti</surname>
<given-names>J. J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Goldsmith</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Riggle</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Yu</surname>
<given-names>A. C. L.</given-names>
</name>
</person-group>
<chapter-title>Intonation</chapter-title>
<source>The handbook of phonological theory</source>
<year iso-8601-date="2011">2011</year>
<publisher-loc>Malden, MA</publisher-loc>
<publisher-name>Blackwell</publisher-name>
<fpage>485</fpage>
<lpage>532</lpage>
<pub-id pub-id-type="doi">10.1002/9781444343069.ch15</pub-id>
</element-citation>
</ref>
<ref id="B28">
<label>28</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Borr&#224;s-Comes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Sichel-Bazin</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>Vocative intonation preferences are sensitive to politeness factors</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2015">2015</year>
<volume>58</volume>
<issue>1</issue>
<fpage>68</fpage>
<lpage>83</lpage>
<pub-id pub-id-type="doi">10.1177/0023830914565441</pub-id>
</element-citation>
</ref>
<ref id="B29">
<label>29</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Borr&#224;s-Comes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Vanrell</surname>
<given-names>M. M.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>The role of pitch range in establishing intonational contrasts</article-title>
<source>Journal of the International Phonetics Association</source>
<year iso-8601-date="2014">2014</year>
<volume>44</volume>
<issue>1</issue>
<fpage>1</fpage>
<lpage>20</lpage>
<pub-id pub-id-type="doi">10.1017/S0025100313000303</pub-id>
</element-citation>
</ref>
<ref id="B30">
<label>30</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Browman</surname>
<given-names>C. P.</given-names>
</name>
<name>
<surname>Goldstein</surname>
<given-names>L.</given-names>
</name>
</person-group>
<article-title>Articulatory phonology: An overview</article-title>
<source>Phonetica</source>
<year iso-8601-date="1992">1992</year>
<volume>49</volume>
<fpage>155</fpage>
<lpage>180</lpage>
<pub-id pub-id-type="doi">10.1159/000261913</pub-id>
</element-citation>
</ref>
<ref id="B31">
<label>31</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Brugos</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Veilleux</surname>
<given-names>N.</given-names>
</name>
</person-group>
<article-title><italic>Transcribing prosodic structure of spoken utterances with ToBI</italic> (MIT Open Courseware)</article-title>
<year iso-8601-date="2006">2006</year>
<comment>Retrieved from <uri>http://ocw.mit.edu/courses/electrical-engineering-and-computer-science/6-911-transcribing-prosodic-structure-of-spoken-utterances-with-tobi-january-iap-2006/index.htm</uri></comment>
</element-citation>
</ref>
<ref id="B32">
<label>32</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Calhoun</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>The centrality of metrical structure in signaling Information Structure: A probabilistic perspective</article-title>
<source>Language</source>
<year iso-8601-date="2010">2010</year>
<volume>86</volume>
<issue>1</issue>
<fpage>1</fpage>
<lpage>42</lpage>
<pub-id pub-id-type="doi">10.1353/lan.0.0197</pub-id>
</element-citation>
</ref>
<ref id="B33">
<label>33</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cangemi</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>The Importance of a Distributional Approach to Categoriality in Autosegmental-Metrical Accounts of Intonation</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>9</elocation-id>
<fpage>1</fpage>
<lpage>20</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.28</pub-id>
</element-citation>
</ref>
<ref id="B34">
<label>34</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Chung</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Speech rhythm in Korean: Experiments in speech cycling</article-title>
<source>Proceedings of Meetings on Acoustics (POMA) 19.060216</source>
<year iso-8601-date="2013">2013</year>
<comment>Retrieved from <uri>http://scitation.aip.org/content/asa/journal/poma/19/1/10.1121/1.4801062</uri></comment>
</element-citation>
</ref>
<ref id="B35">
<label>35</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Clopper</surname>
<given-names>C. G.</given-names>
</name>
<name>
<surname>Smiljanic</surname>
<given-names>R.</given-names>
</name>
</person-group>
<article-title>Effects of gender and regional dialect on prosodic patterns in American English</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2011">2011</year>
<volume>39</volume>
<issue>2</issue>
<fpage>237</fpage>
<lpage>245</lpage>
<pub-id pub-id-type="doi">10.1016/j.wocn.2011.02.006</pub-id>
</element-citation>
</ref>
<ref id="B36">
<label>36</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cole</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>New Methods for Prosodic Transcription: Capturing Variability as a Source of Information</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>8</elocation-id>
<fpage>1</fpage>
<lpage>29</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.29</pub-id>
</element-citation>
</ref>
<ref id="B37">
<label>37</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Cruttenden</surname>
<given-names>A.</given-names>
</name>
</person-group>
<source>Gimson&#8217;s pronunciation of English</source>
<year iso-8601-date="1994">1994</year>
<publisher-loc>London</publisher-loc>
<publisher-name>Edward Arnold</publisher-name>
</element-citation>
</ref>
<ref id="B38">
<label>38</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Crystal</surname>
<given-names>D.</given-names>
</name>
</person-group>
<chapter-title>Review of M. A. K. Halliday, <italic>Intonation and Grammar in British English</italic></chapter-title>
<source>Language</source>
<year iso-8601-date="1969">1969</year>
<publisher-loc>The Hague</publisher-loc>
<publisher-name>Mouton</publisher-name>
<volume>45</volume>
<issue>2</issue>
<fpage>378</fpage>
<lpage>393</lpage>
<pub-id pub-id-type="doi">10.2307/411669</pub-id>
<comment>1967</comment>
</element-citation>
</ref>
<ref id="B39">
<label>39</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cummins</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Port</surname>
<given-names>R. F.</given-names>
</name>
</person-group>
<article-title>Rhythmic constraints on stress-timing in English</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="1998">1998</year>
<volume>31</volume>
<fpage>139</fpage>
<lpage>148</lpage>
<pub-id pub-id-type="doi">10.1016/S0095-4470(02)00082-7</pub-id>
</element-citation>
</ref>
<ref id="B40">
<label>40</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>D&#8217;Imperio</surname>
<given-names>M.</given-names>
</name>
</person-group>
<source>The role of perception in defining tonal targets and their alignment. (Unpublished doctoral dissertation)</source>
<year iso-8601-date="2000">2000</year>
<publisher-name>The Ohio State University</publisher-name>
</element-citation>
</ref>
<ref id="B41">
<label>41</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>D&#8217;Imperio</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Terken</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Piterman</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Perceived tone &#8220;targets&#8221; and pitch accent identification in Italian</article-title>
<conf-name>Proceedings of Australian International Conference on Speech Science and Technology (SST)</conf-name>
<year iso-8601-date="2000">2000</year>
<volume>8</volume>
<fpage>201</fpage>
<lpage>211</lpage>
</element-citation>
</ref>
<ref id="B42">
<label>42</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Edwards</surname>
<given-names>J. R.</given-names>
</name>
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Munson</surname>
<given-names>B.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Redford</surname>
<given-names>M.</given-names>
</name>
</person-group>
<chapter-title>Cross-language differences in acquisition</chapter-title>
<source>The handbook of speech production</source>
<year iso-8601-date="2015">2015</year>
<publisher-loc>Malden, MA</publisher-loc>
<publisher-name>Wiley-Blackwell</publisher-name>
<fpage>530</fpage>
<lpage>554</lpage>
</element-citation>
</ref>
<ref id="B43">
<label>43</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Fougeron</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<article-title>Rate effects on French intonation: Phonetic realization and prosodic organization</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="1998">1998</year>
<volume>26</volume>
<issue>1</issue>
<fpage>45</fpage>
<lpage>70</lpage>
<pub-id pub-id-type="doi">10.1006/jpho.1997.0062</pub-id>
</element-citation>
</ref>
<ref id="B44">
<label>44</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<collab>FRA &amp; UNDP: European Union Agency for Fundamental Rights United Nations Development Program</collab>
</person-group>
<article-title>The situation of Roma in 11 EU states: Survey results at a glance. Luxembourg: Publications Office of the European Union</article-title>
<year iso-8601-date="2012">2012</year>
<comment>Retrieved from <uri>http://www.scribd.com/doc/153872420/The-situation-of-Roma-in-11-EU-Member-States</uri></comment>
</element-citation>
</ref>
<ref id="B45">
<label>45</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Frota</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Surface and Structure: Transcribing Intonation within and across Languages</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>7</elocation-id>
<fpage>1</fpage>
<lpage>19</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.10</pub-id>
<comment>5</comment>
</element-citation>
</ref>
<ref id="B46">
<label>46</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Frota</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Cruz</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Svartman</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Vig&#225;rio</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Collischonn</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Fonseca</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Serra</surname>
<given-names>C.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Frota</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<chapter-title>Intonational variation in Portuguese: European and Brazilian varieties</chapter-title>
<source>Intonation variation in Romance</source>
<year iso-8601-date="2015">2015</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>235</fpage>
<lpage>283</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199685332.003.0007</pub-id>
<uri>http://dx.doi.org/10.1093/acprof:oso/9780199685332.001.0001</uri>
</element-citation>
</ref>
<ref id="B47">
<label>47</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Fry</surname>
<given-names>D. B.</given-names>
</name>
</person-group>
<article-title>Experiments in the perception of stress</article-title>
<source>Language and Speech</source>
<year iso-8601-date="1958">1958</year>
<volume>1</volume>
<fpage>126</fpage>
<lpage>152</lpage>
</element-citation>
</ref>
<ref id="B48">
<label>48</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>Grabe</surname>
<given-names>E.</given-names>
</name>
</person-group>
<source>Intonational phonology: English and German (Unpublished doctoral dissertation)</source>
<year iso-8601-date="1998">1998</year>
<publisher-name>Max-Planck-Institute for Psycholinguistics and University of Nijmegen</publisher-name>
<comment>Retrieved from <uri>http://www.phon.ox.ac.uk/files/people/grabe/thesis.html</uri></comment>
</element-citation>
</ref>
<ref id="B49">
<label>49</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Baumann</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Benzm&#252;ller</surname>
<given-names>R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>German intonation in autosegmental-metrical phonology</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005">2005</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>55</fpage>
<lpage>83</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0003</pub-id>
</element-citation>
</ref>
<ref id="B50">
<label>50</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>On the place of &#8220;phrase accents&#8221; in intonational phonology</article-title>
<source>Phonology</source>
<year iso-8601-date="2000">2000</year>
<volume>17</volume>
<fpage>143</fpage>
<lpage>185</lpage>
<pub-id pub-id-type="doi">10.1017/S0952675700003924</pub-id>
</element-citation>
</ref>
<ref id="B51">
<label>51</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<source>On the grammar and semantics of sentence accents</source>
<year iso-8601-date="1984">1984</year>
<publisher-loc>Dordrecht</publisher-loc>
<publisher-name>Foris</publisher-name>
<pub-id pub-id-type="doi">10.1515/9783110859263</pub-id>
</element-citation>
</ref>
<ref id="B52">
<label>52</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<source>The phonology of tone and intonation</source>
<year iso-8601-date="2004">2004</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511616983</pub-id>
</element-citation>
</ref>
<ref id="B53">
<label>53</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>Transcription of Dutch intonation</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005">2005</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>118</fpage>
<lpage>145</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0005</pub-id>
</element-citation>
</ref>
<ref id="B54">
<label>54</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>de Lacy</surname>
<given-names>P.</given-names>
</name>
</person-group>
<chapter-title>The phonology of intonation</chapter-title>
<source>The Cambridge handbook of phonology</source>
<year iso-8601-date="2007">2007</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<fpage>253</fpage>
<lpage>280</lpage>
</element-citation>
</ref>
<ref id="B55">
<label>55</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<article-title>Analysis of Intonation: the Case of MAE_ToBI</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>10</elocation-id>
<fpage>1</fpage>
<lpage>35</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.30</pub-id>
</element-citation>
</ref>
<ref id="B56">
<label>56</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Halliday</surname>
<given-names>M. A. K.</given-names>
</name>
</person-group>
<source>Intonation and grammar in British English</source>
<year iso-8601-date="1967">1967</year>
<publisher-loc>The Hague</publisher-loc>
<publisher-name>Mouton</publisher-name>
<pub-id pub-id-type="doi">10.1515/9783111357447</pub-id>
</element-citation>
</ref>
<ref id="B57">
<label>57</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Haspelmath</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Comparative concepts and descriptive categories in crosslinguistic studies</article-title>
<source>Language</source>
<year iso-8601-date="2010">2010</year>
<volume>86</volume>
<issue>3</issue>
<fpage>663</fpage>
<lpage>687</lpage>
<pub-id pub-id-type="doi">10.1353/lan.2010.0021</pub-id>
</element-citation>
</ref>
<ref id="B58">
<label>58</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Haspelmath</surname>
<given-names>M.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>B&#322;aszczak</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Klimek-Jankowska</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Migdalski</surname>
<given-names>K.</given-names>
</name>
</person-group>
<chapter-title>Defining vs. diagnosing linguistic categories: a case study of clitic phenomena</chapter-title>
<source>How categorical are categories?</source>
<year iso-8601-date="2015">2015</year>
<publisher-loc>Berlin</publisher-loc>
<publisher-name>De Gruyter Mouton</publisher-name>
<fpage>273</fpage>
<lpage>304</lpage>
<pub-id pub-id-type="doi">10.1515/9781614514510-009</pub-id>
</element-citation>
</ref>
<ref id="B59">
<label>59</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Henrich</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Heine</surname>
<given-names>S. J.</given-names>
</name>
<name>
<surname>Norenzayan</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>The weirdest people in the world?</article-title>
<source>Behavioral and Brain Sciences</source>
<year iso-8601-date="2010">2010</year>
<volume>33</volume>
<fpage>61</fpage>
<lpage>135</lpage>
<pub-id pub-id-type="doi">10.1017/S0140525X0999152X</pub-id>
</element-citation>
</ref>
<ref id="B60">
<label>60</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Brown</surname>
<given-names>K.</given-names>
</name>
</person-group>
<chapter-title>Speech synthesis: prosody</chapter-title>
<source>Encyclopedia of language and linguistics</source>
<year iso-8601-date="2006">2006</year>
<edition>2nd ed.</edition>
<publisher-loc>Amsterdam</publisher-loc>
<publisher-name>Elsevier</publisher-name>
<fpage>49</fpage>
<lpage>55</lpage>
<pub-id pub-id-type="doi">10.1016/B0-08-044854-2/00914-7</pub-id>
</element-citation>
</ref>
<ref id="B61">
<label>61</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Hualde</surname>
<given-names>J. I.</given-names>
</name>
<name>
<surname>Elordieta</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Gaminde</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Smiljani&#263;</surname>
<given-names>R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Warner</surname>
<given-names>N.</given-names>
</name>
</person-group>
<chapter-title>From pitch accent to stress accent in Basque</chapter-title>
<source>Laboratory Phonology 7</source>
<year iso-8601-date="2002">2002</year>
<publisher-loc>Berlin, New York</publisher-loc>
<publisher-name>Mouton de Gruyter</publisher-name>
<fpage>547</fpage>
<lpage>584</lpage>
<pub-id pub-id-type="doi">10.1515/9783110197105.547</pub-id>
</element-citation>
</ref>
<ref id="B62">
<label>62</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hualde</surname>
<given-names>J. I.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>Towards an International Prosodic Alphabet (IPrA)</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>5</elocation-id>
<fpage>1</fpage>
<lpage>25</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.11</pub-id>
</element-citation>
</ref>
<ref id="B63">
<label>63</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hyman</surname>
<given-names>L. M.</given-names>
</name>
</person-group>
<article-title>Word-prosodic typology</article-title>
<source>Phonology</source>
<year iso-8601-date="2006">2006</year>
<volume>23</volume>
<fpage>225</fpage>
<lpage>257</lpage>
<pub-id pub-id-type="doi">10.1017/S0952675706000893</pub-id>
</element-citation>
</ref>
<ref id="B64">
<label>64</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<collab>IPA</collab>
</person-group>
<source>Handbook of the International Phonetic Association</source>
<year iso-8601-date="1999">1999</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
</element-citation>
</ref>
<ref id="B65">
<label>65</label>
<element-citation publication-type="book">
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005a">2005a</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.001.0001</pub-id>
</element-citation>
</ref>
<ref id="B66">
<label>66</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>Korean intonational phonology and prosodic transcription</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005b">2005b</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>201</fpage>
<lpage>229</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0008</pub-id>
<uri>http://dx.doi.org/10.1093/acprof:oso/9780199249633.001.0001</uri>
</element-citation>
</ref>
<ref id="B67">
<label>67</label>
<element-citation publication-type="book">
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>Prosodic typology II: The phonology of intonation and phrasing</chapter-title>
<year iso-8601-date="2014">2014</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199567300.001.0001</pub-id>
</element-citation>
</ref>
<ref id="B68">
<label>68</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
<name>
<surname>Fletcher</surname>
<given-names>J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>Methodology of studying intonation: from data collection to data analysis</chapter-title>
<source>Prosodic typology II: The phonology of intonation and phrasing</source>
<year iso-8601-date="2014">2014</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>493</fpage>
<lpage>519</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199567300.003.0016</pub-id>
<uri>http://dx.doi.org/10.1093/acprof:oso/9780199567300.001.0001</uri>
</element-citation>
</ref>
<ref id="B69">
<label>69</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Keating</surname>
<given-names>P. A.</given-names>
</name>
</person-group>
<article-title>Phonetic and phonological representation of stop consonant voicing</article-title>
<source>Language</source>
<year iso-8601-date="1984">1984</year>
<volume>60</volume>
<issue>2</issue>
<fpage>285</fpage>
<lpage>319</lpage>
<pub-id pub-id-type="doi">10.2307/413642</pub-id>
</element-citation>
</ref>
<ref id="B70">
<label>70</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Kim</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<article-title>Prosodic structure and focus prosody of South Kyungsang Korean</article-title>
<source>Language Research</source>
<year iso-8601-date="2009">2009</year>
<volume>45</volume>
<issue>1</issue>
<fpage>43</fpage>
<lpage>66</lpage>
</element-citation>
</ref>
<ref id="B71">
<label>71</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Knight</surname>
<given-names>R.-A.</given-names>
</name>
</person-group>
<article-title>The shape of nuclear falls and their effect on the perception of pitch and prominence: peaks vs. plateaux</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2008">2008</year>
<volume>51</volume>
<issue>3</issue>
<fpage>223</fpage>
<lpage>244</lpage>
<pub-id pub-id-type="doi">10.1177/0023830908098541</pub-id>
</element-citation>
</ref>
<ref id="B72">
<label>72</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Knight</surname>
<given-names>R.-A.</given-names>
</name>
<name>
<surname>Nolan</surname>
<given-names>F.</given-names>
</name>
</person-group>
<article-title>The effect of pitch span on intonational plateaux</article-title>
<source>Journal of the International Phonetic Association</source>
<year iso-8601-date="2006">2006</year>
<volume>36</volume>
<issue>1</issue>
<fpage>21</fpage>
<lpage>38</lpage>
<pub-id pub-id-type="doi">10.1017/S0025100306002349</pub-id>
</element-citation>
</ref>
<ref id="B73">
<label>73</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Kochanski</surname>
<given-names>G. J.</given-names>
</name>
</person-group>
<article-title>Prosodic peak estimation under segmental perturbations</article-title>
<source>Journal of the Acoustical Society of America</source>
<year iso-8601-date="2010">2010</year>
<volume>127</volume>
<issue>2</issue>
<fpage>862</fpage>
<lpage>873</lpage>
<pub-id pub-id-type="doi">10.1121/1.3268511</pub-id>
</element-citation>
</ref>
<ref id="B74">
<label>74</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<article-title>Phonological features of intonational peaks</article-title>
<source>Language</source>
<year iso-8601-date="1983">1983</year>
<volume>59</volume>
<fpage>721</fpage>
<lpage>759</lpage>
<pub-id pub-id-type="doi">10.2307/413371</pub-id>
</element-citation>
</ref>
<ref id="B75">
<label>75</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<article-title>Perfect pitch</article-title>
<source>New Scientist</source>
<year iso-8601-date="1999">1999</year>
<volume>164</volume>
<issue>2218</issue>
<fpage>87</fpage>
</element-citation>
</ref>
<ref id="B76">
<label>76</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<source>Intonational phonology</source>
<year iso-8601-date="2008a">2008a</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511808814</pub-id>
</element-citation>
</ref>
<ref id="B77">
<label>77</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>Sun-Ah</given-names>
</name>
</person-group>
<chapter-title>Review of Prosodic typology: the phonology of intonation and phrasing</chapter-title>
<source>Phonology</source>
<year iso-8601-date="2008b">2008b</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<volume>25</volume>
<fpage>372</fpage>
<lpage>376</lpage>
<comment>(2005)</comment>
</element-citation>
</ref>
<ref id="B78">
<label>78</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Goldsmith</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Riggle</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Yu</surname>
<given-names>A. C. L.</given-names>
</name>
</person-group>
<chapter-title>Phonetics in phonology</chapter-title>
<source>Handbook of phonological theory</source>
<year iso-8601-date="2011">2011</year>
<publisher-loc>Malden, MA</publisher-loc>
<publisher-name>Blackwell</publisher-name>
<fpage>348</fpage>
<lpage>373</lpage>
<pub-id pub-id-type="doi">10.1002/9781444343069.ch11</pub-id>
</element-citation>
</ref>
<ref id="B79">
<label>79</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Morton</surname>
<given-names>R.</given-names>
</name>
</person-group>
<article-title>The perception of intonational emphasis: continuous or categorical?</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="1997">1997</year>
<volume>25</volume>
<fpage>313</fpage>
<lpage>342</lpage>
<pub-id pub-id-type="doi">10.1006/jpho.1997.0046</pub-id>
</element-citation>
</ref>
<ref id="B80">
<label>80</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Schepman</surname>
<given-names>A.</given-names>
</name>
</person-group>
<chapter-title>&#8220;Sagging transitions&#8221; between high accent peaks in English: &#8220;experimental evidence&#8221;</chapter-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2003">2003</year>
<volume>31</volume>
<fpage>81</fpage>
<lpage>112</lpage>
<pub-id pub-id-type="doi">10.1016/S0095-4470(02)00073-6</pub-id>
</element-citation>
</ref>
<ref id="B81">
<label>81</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladefoged</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Johnson</surname>
<given-names>K.</given-names>
</name>
</person-group>
<source>A course in phonetics</source>
<year iso-8601-date="2011">2011</year>
<publisher-loc>Boston</publisher-loc>
<publisher-name>Wadsworth Cengage Learning</publisher-name>
</element-citation>
</ref>
<ref id="B82">
<label>82</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>Liberman</surname>
<given-names>M.</given-names>
</name>
</person-group>
<source>The intonational system of English (Unpublished doctoral thesis)</source>
<year iso-8601-date="1975">1975</year>
<publisher-name>MIT</publisher-name>
</element-citation>
</ref>
<ref id="B83">
<label>83</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Martins-Heub</surname>
<given-names>K.</given-names>
</name>
</person-group>
<article-title>Genocide in the 20th century: Reflections on the collective identify of German Roma and Sinti (Gypsies) after national socialism</article-title>
<source>Holocaust Genocide Studies</source>
<year iso-8601-date="1989">1989</year>
<volume>4</volume>
<issue>2</issue>
<fpage>193</fpage>
<lpage>211</lpage>
<pub-id pub-id-type="doi">10.1093/hgs/4.2.193</pub-id>
</element-citation>
</ref>
<ref id="B84">
<label>84</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Matras</surname>
<given-names>Y.</given-names>
</name>
</person-group>
<article-title>Writing Romani: The pragmatics of codification in a stateless language</article-title>
<source>Applied Linguistics</source>
<year iso-8601-date="1999">1999</year>
<volume>20</volume>
<issue>4</issue>
<fpage>481</fpage>
<lpage>502</lpage>
<pub-id pub-id-type="doi">10.1093/applin/20.4.481</pub-id>
</element-citation>
</ref>
<ref id="B85">
<label>85</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Matras</surname>
<given-names>Y.</given-names>
</name>
</person-group>
<source>Romani: A linguistic introduction</source>
<year iso-8601-date="2002">2002</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511486791</pub-id>
</element-citation>
</ref>
<ref id="B86">
<label>86</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Matras</surname>
<given-names>Y.</given-names>
</name>
</person-group>
<article-title>The future of Romani: Toward a policy of linguistic pluralism</article-title>
<source>Roma Rights Quarterly</source>
<year iso-8601-date="2005">2005</year>
<volume>1</volume>
<fpage>31</fpage>
<lpage>44</lpage>
</element-citation>
</ref>
<ref id="B87">
<label>87</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Mixdorff</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Leemann</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Dellwo</surname>
<given-names>V.</given-names>
</name>
</person-group>
<article-title>The influence of speech rate on Fujisaki model parameters</article-title>
<source>EURASIP Journal on Audio, Speech, and Music Processing</source>
<year iso-8601-date="2014">2014</year>
<volume>33</volume>
<pub-id pub-id-type="doi">10.1186/s13636-014-0033-6</pub-id>
<comment>Retrieved from <uri>http://asmp.eurasipjournals.com/content/2014/1/33</uri></comment>
</element-citation>
</ref>
<ref id="B88">
<label>88</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Niebuhr</surname>
<given-names>O.</given-names>
</name>
</person-group>
<article-title>Coding of intonational meanings beyond F0: Evidence from utterance-final /t/ aspiration in German</article-title>
<source>Journal of the Acoustical Society of America</source>
<year iso-8601-date="2008">2008</year>
<volume>124</volume>
<issue>2</issue>
<fpage>1252</fpage>
<lpage>1263</lpage>
<pub-id pub-id-type="doi">10.1121/1.2940588</pub-id>
</element-citation>
</ref>
<ref id="B89">
<label>89</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Niebuhr</surname>
<given-names>O.</given-names>
</name>
</person-group>
<chapter-title>At the edge of intonation &#8211; The interplay of utterance-final F0 movements and voiceless fricative sounds in German</chapter-title>
<source>Phonetica</source>
<year iso-8601-date="2012">2012</year>
<volume>69</volume>
<issue>1&#8211;2</issue>
<fpage>7</fpage>
<lpage>21</lpage>
<pub-id pub-id-type="doi">10.1159/000343171</pub-id>
</element-citation>
</ref>
<ref id="B90">
<label>90</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>O&#8217;Connor</surname>
<given-names>J. D.</given-names>
</name>
<name>
<surname>Arnold</surname>
<given-names>G. F.</given-names>
</name>
</person-group>
<source>Intonation of colloquial English: A practical handbook</source>
<year iso-8601-date="1973">1973</year>
<publisher-loc>London</publisher-loc>
<publisher-name>Longman</publisher-name>
</element-citation>
</ref>
<ref id="B91">
<label>91</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Peters</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Hanssen</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<article-title>The timing of nuclear falls: Evidence from Dutch, West Frisian, Dutch Low Saxon, German Low Saxon, and High German</article-title>
<source>Laboratory Phonology</source>
<year iso-8601-date="2015">2015</year>
<volume>6</volume>
<issue>1</issue>
<fpage>1</fpage>
<lpage>52</lpage>
<pub-id pub-id-type="doi">10.1515/lp-2015-0004</pub-id>
</element-citation>
</ref>
<ref id="B92">
<label>92</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J.</given-names>
</name>
</person-group>
<source>The phonology and phonetics of English intonation (Unpublished doctoral thesis)</source>
<year iso-8601-date="1980">1980</year>
<publisher-name>MIT</publisher-name>
</element-citation>
</ref>
<ref id="B93">
<label>93</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Warner</surname>
<given-names>N.</given-names>
</name>
</person-group>
<chapter-title>Word-specific phonetics</chapter-title>
<source>Laboratory Phonology VII</source>
<year iso-8601-date="2002">2002</year>
<publisher-loc>Berlin, New York</publisher-loc>
<publisher-name>Mouton de Gruyter</publisher-name>
<fpage>101</fpage>
<lpage>139</lpage>
<pub-id pub-id-type="doi">10.1515/9783110197105.101</pub-id>
</element-citation>
</ref>
<ref id="B94">
<label>94</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Beckman</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Burton-Roberts</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Carr</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Docherty</surname>
<given-names>G. J.</given-names>
</name>
</person-group>
<chapter-title>Conceptual foundations of phonology as a laboratory science</chapter-title>
<source>Conceptual and empirical foundations of phonology</source>
<year iso-8601-date="2000">2000</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>273</fpage>
<lpage>303</lpage>
<comment>Reprinted 2012 in Cohn, A., Fougeron, C., &amp; Huffman, M. (Eds.), The Oxford handbook of laboratory phonology. Oxford:: Oxford University Press. pp. 17&#8211;39</comment>
</element-citation>
</ref>
<ref id="B95">
<label>94</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Cohen</surname>
<given-names>P. R.</given-names>
</name>
<name>
<surname>Morgan</surname>
<given-names>J. L.</given-names>
</name>
<name>
<surname>Pollack</surname>
<given-names>M. E.</given-names>
</name>
</person-group>
<chapter-title>The meaning of intonational contours in the interpretation of discourse</chapter-title>
<source>Intentions in communication</source>
<year iso-8601-date="1990">1990</year>
<publisher-loc>Cambridge MA</publisher-loc>
<publisher-name>The MIT Press</publisher-name>
<fpage>271</fpage>
<lpage>311</lpage>
</element-citation>
</ref>
<ref id="B96">
<label>96</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Podesva</surname>
<given-names>R. J.</given-names>
</name>
</person-group>
<article-title>Phonation type as a stylistic variable: The use of falsetto in constructing a persona</article-title>
<source>Journal of Sociolinguistics</source>
<year iso-8601-date="2007">2007</year>
<volume>11</volume>
<fpage>478</fpage>
<lpage>504</lpage>
<pub-id pub-id-type="doi">10.1111/j.1467-9841.2007.00334.x</pub-id>
</element-citation>
</ref>
<ref id="B97">
<label>97</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Port</surname>
<given-names>R. F.</given-names>
</name>
<name>
<surname>Leary</surname>
<given-names>A. P.</given-names>
</name>
</person-group>
<article-title>Against formal phonology</article-title>
<source>Language</source>
<year iso-8601-date="2005">2005</year>
<volume>81</volume>
<issue>4</issue>
<fpage>927</fpage>
<lpage>964</lpage>
<pub-id pub-id-type="doi">10.1353/lan.2005.0195</pub-id>
</element-citation>
</ref>
<ref id="B98">
<label>98</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>The scaling of the L tone line in Spanish downstepping contours</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="1998">1998</year>
<volume>26</volume>
<fpage>261</fpage>
<lpage>282</lpage>
<pub-id pub-id-type="doi">10.1006/jpho.1998.0074</pub-id>
</element-citation>
</ref>
<ref id="B99">
<label>99</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>Word-edge tones in Catalan</article-title>
<source>Italian Journal of Linguistics</source>
<year iso-8601-date="2006">2006</year>
<volume>18</volume>
<issue>1</issue>
<fpage>39</fpage>
<lpage>71</lpage>
</element-citation>
</ref>
<ref id="B100">
<label>100</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>D&#8217;Imperio</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Gili-Fivela</surname>
<given-names>B.</given-names>
</name>
</person-group>
<article-title>Pitch accent alignment in Romance: primary and secondary associations with metrical structure</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2005">2005</year>
<volume>48</volume>
<issue>4</issue>
<fpage>359</fpage>
<lpage>396</lpage>
<pub-id pub-id-type="doi">10.1177/00238309050480040301</pub-id>
</element-citation>
</ref>
<ref id="B101">
<label>101</label>
<element-citation publication-type="book">
<person-group person-group-type="editor">
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Roseano</surname>
<given-names>P.</given-names>
</name>
</person-group>
<source>Transcription of intonation of the Spanish language</source>
<year iso-8601-date="2010">2010</year>
<publisher-loc>Lincom Europa</publisher-loc>
<publisher-name>M&#252;nchen</publisher-name>
</element-citation>
</ref>
<ref id="B102">
<label>102</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Ritchart</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>The form and use of uptalk in Southern California English</article-title>
<conf-name>Proceedings of Speech Prosody 7</conf-name>
<conf-date>May 20&#8211;23, 2014</conf-date>
<year iso-8601-date="2014">2014</year>
<conf-loc>Dublin</conf-loc>
<comment>Retrieved from <uri>http://www.speechprosody2014.org</uri></comment>
</element-citation>
</ref>
<ref id="B103">
<label>103</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ritter</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>The role of tonal onglides in German nuclear pitch accents</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2015">2015</year>
<volume>58</volume>
<issue>1</issue>
<fpage>114</fpage>
<lpage>128</lpage>
<pub-id pub-id-type="doi">10.1177/0023830914565688</pub-id>
</element-citation>
</ref>
<ref id="B104">
<label>104</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Sachs</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Schegloff</surname>
<given-names>E. A.</given-names>
</name>
<name>
<surname>Jefferson</surname>
<given-names>G.</given-names>
</name>
</person-group>
<article-title>A simplest systematics for the organization of turn-taking in conversation</article-title>
<source>Language</source>
<year iso-8601-date="1974">1974</year>
<volume>50</volume>
<fpage>696</fpage>
<lpage>735</lpage>
<pub-id pub-id-type="doi">10.2307/412243</pub-id>
<uri>http://dx.doi.org/10.1353/lan.1974.0010</uri>
</element-citation>
</ref>
<ref id="B105">
<label>105</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Scobbie</surname>
<given-names>J. M.</given-names>
</name>
<name>
<surname>Gibbon</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Hardcastle</surname>
<given-names>W. J.</given-names>
</name>
<name>
<surname>Fletcher</surname>
<given-names>P.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Broe</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J.</given-names>
</name>
</person-group>
<chapter-title>Covert contrast as a stage in the acquisition of phonetics and phonology</chapter-title>
<source>Papers in Laboratory Phonology V: Language Acquisition and the Lexicon</source>
<year iso-8601-date="2000">2000</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<fpage>194</fpage>
<lpage>207</lpage>
</element-citation>
</ref>
<ref id="B106">
<label>106</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Silverman</surname>
<given-names>K. E. A.</given-names>
</name>
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Pitrelli</surname>
<given-names>J. F.</given-names>
</name>
<name>
<surname>Ostendorf</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Wightman</surname>
<given-names>C. W.</given-names>
</name>
<name>
<surname>Price</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>ToBI: A standard for labeling English prosody</article-title>
<conf-name>Proceedings of ICSLP 1992</conf-name>
<year iso-8601-date="1992">1992</year>
</element-citation>
</ref>
<ref id="B107">
<label>107</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Skopeteas</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Fiedler</surname>
<given-names>I.</given-names>
</name>
<name>
<surname>Hellmuth</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Schwarz</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Stoel</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Fanselow</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>F&#233;ry</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Krifka</surname>
<given-names>M.</given-names>
</name>
</person-group>
<source>Questionnaire on information structure</source>
<year iso-8601-date="2006">2006</year>
<publisher-loc>Potsdam</publisher-loc>
<publisher-name>Universit&#228;t Potsdam, Institut f&#252;r Linguistik/Allgemeine Sprachwissenschaft</publisher-name>
</element-citation>
</ref>
<ref id="B108">
<label>108</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Sproat</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Hunt</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ostendorf</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Taylor</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Black</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Lenzo</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Edgington</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>SABLE: A standard for TTS markup</article-title>
<conf-name>International Conference on Spoken Language Processing, 1998</conf-name>
<year iso-8601-date="1998">1998</year>
</element-citation>
</ref>
<ref id="B109">
<label>109</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Stuart-Smith</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Sonderegger</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Rathcke</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Macdonald</surname>
<given-names>R.</given-names>
</name>
</person-group>
<article-title>The private life of stops: VOT in a real-time corpus of spontaneous Glaswegian</article-title>
<source>Laboratory Phonology</source>
<year iso-8601-date="2015">2015</year>
<volume>6</volume>
<issue>3&#8211;4</issue>
<fpage>505</fpage>
<lpage>549</lpage>
<pub-id pub-id-type="doi">10.1515/lp-2015-0015</pub-id>
</element-citation>
</ref>
<ref id="B110">
<label>110</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Swerts</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Krahmer</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Avesani</surname>
<given-names>C.</given-names>
</name>
</person-group>
<article-title>Prosodic marking of information status in Dutch and Italian: A comparative analysis</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2002">2002</year>
<volume>30</volume>
<fpage>629</fpage>
<lpage>654</lpage>
<pub-id pub-id-type="doi">10.1006/jpho.2002.0178</pub-id>
</element-citation>
</ref>
<ref id="B111">
<label>111</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Tannen</surname>
<given-names>D.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Dechert</surname>
<given-names>H.-W.</given-names>
</name>
<name>
<surname>Raupach</surname>
<given-names>M.</given-names>
</name>
</person-group>
<chapter-title>Conversational style</chapter-title>
<source>Psycholinguistic models of production</source>
<year iso-8601-date="1987">1987</year>
<publisher-loc>Norwood, NJ</publisher-loc>
<publisher-name>Ablex</publisher-name>
<fpage>251</fpage>
<lpage>267</lpage>
</element-citation>
</ref>
<ref id="B112">
<label>112</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Tyaglyy</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Were the &#8220;Chingen&#233;&#8221; victims of the Holocaust? Nazi policy toward the Crimean Roma, 1941&#8211;1944</article-title>
<source>Holocaust Genocide Studies</source>
<year iso-8601-date="2009">2009</year>
<volume>23</volume>
<issue>1</issue>
<fpage>26</fpage>
<lpage>53</lpage>
<pub-id pub-id-type="doi">10.1093/hgs/dcp015</pub-id>
</element-citation>
</ref>
<ref id="B113">
<label>113</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>van Santen</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Shih</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>M&#246;bius</surname>
<given-names>B.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Sproat</surname>
<given-names>R.</given-names>
</name>
</person-group>
<chapter-title>Intonation</chapter-title>
<source>Multilingual text-to-speech synthesis: The Bell Labs approach</source>
<year iso-8601-date="1997">1997</year>
<publisher-loc>Dordrecht, Boston, London</publisher-loc>
<publisher-name>Kluwer Academic Publishers</publisher-name>
<fpage>141</fpage>
<lpage>189</lpage>
</element-citation>
</ref>
<ref id="B114">
<label>114</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Varga</surname>
<given-names>L.</given-names>
</name>
</person-group>
<article-title>The calling contour in Hungarian and English</article-title>
<source>Phonology</source>
<year iso-8601-date="2008">2008</year>
<volume>25</volume>
<fpage>469</fpage>
<lpage>497</lpage>
<pub-id pub-id-type="doi">10.1017/S0952675708001607</pub-id>
</element-citation>
</ref>
<ref id="B115">
<label>115</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Venditti</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Maekawa</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Miyagawa</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Saito</surname>
<given-names>M.</given-names>
</name>
</person-group>
<chapter-title>Prominence marking in the Japanese intonation system</chapter-title>
<source>Handbook of Japanese linguistics</source>
<year iso-8601-date="2008">2008</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>456</fpage>
<lpage>512</lpage>
<pub-id pub-id-type="doi">10.1093/oxfordhb/9780195307344.013.0017</pub-id>
</element-citation>
</ref>
<ref id="B116">
<label>116</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Wells</surname>
<given-names>J. C.</given-names>
</name>
</person-group>
<source>Accents of English</source>
<year iso-8601-date="1982">1982</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511611759</pub-id>
</element-citation>
</ref>
<ref id="B117">
<label>117</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Xu</surname>
<given-names>Y.</given-names>
</name>
</person-group>
<article-title>In defense of lab speech</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2010">2010</year>
<volume>38</volume>
<fpage>329</fpage>
<lpage>336</lpage>
<pub-id pub-id-type="doi">10.1016/j.wocn.2010.04.003</pub-id>
</element-citation>
</ref>
<ref id="B118">
<label>118</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Yuasa</surname>
<given-names>I. P.</given-names>
</name>
</person-group>
<article-title>Creaky voice: A new feminine voice quality for young urban-oriented upwardly mobile American women?</article-title>
<source>American Speech</source>
<year iso-8601-date="2010">2010</year>
<volume>8</volume>
<issue>3</issue>
<fpage>315</fpage>
<lpage>337</lpage>
<pub-id pub-id-type="doi">10.1215/00031283-2010-018</pub-id>
</element-citation>
</ref>
</ref-list>
</back>
</article>