<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.0 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.0/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.0" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Ubiquity Press</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.5334/labphon.30</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Analysis of Intonation: the Case of MAE_ToBI</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>Carlos</given-names>
</name>
<email>c.gussenhoven@let.ru.nl</email>
<xref ref-type="aff" rid="aff-1"/>
</contrib>
</contrib-group>
<aff id="aff-1">Radboud Universiteit Nijmegen, Netherlands</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2016-06-30">
<day>30</day>
<month>06</month>
<year>2016</year>
</pub-date>
<volume>7</volume>
<issue>1</issue>
<elocation-id>10</elocation-id>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2016 The Author(s)</copyright-statement>
<copyright-year>2016</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.5334/labphon.30/"/>
<abstract>
<p>Annotation systems for intonation contours are ideally based on a well-motivated phonological analysis of the language in question, such that instances of indecision are restricted to uncertainties over what intonational structure the speaker has used, rather than over the choice of label in situations where no suitably distinctive label is available or more than one suitable label is available. This contribution inventorizes a number of cases of overanalysis and underanalysis in MAE_ToBI and argues that they are in large part due to the decision by Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) to analyze a rising-falling accent as a rising pitch accent (L+H*) followed by a L-tone from a different source (an &#8216;on-ramp&#8217; analysis). It is shown how the opposite choice, a falling pitch accent preceded by a L-tone from a different source (an &#8216;off-ramp&#8217; analysis), avoids most of these problems. Results from a perception experiment testing MAE_ToBI&#8217;s prediction of intonational boundaries show that steep falls do not always signal a boundary. The inclusion of a tritonal prenuclear pitch accent, which explains the absence of an intonational boundary after a steep fall followed by a gradual rise, can readily be accommodated in the &#8216;off-ramp&#8217; analysis, but not in MAE_ToBI.</p>
</abstract>
<kwd-group>
<kwd>ToBI</kwd>
<kwd>intonational annotation</kwd>
<kwd>intonational phonology</kwd>
<kwd>English</kwd>
<kwd>intonational phrasing</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>Publisher&#039;s Note</title>
<p>A correction article relating to this publication can be found here: <uri>http://dx.doi.org/10.5334/labphon.60</uri>.
</p>
</sec>
<sec>
<title>1 Introduction</title>
<p>Faced with the need to identify the phonological elements in a single rising-falling accent peak in an otherwise low-pitched intonation contour, an analyst has three options, all of which have been adopted for West Germanic (Figure <xref ref-type="fig" rid="F1">1</xref>). First, the rise could be the pitch accent and the fall a transition between H and a following L-tone. This option was taken by Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>), as L+H*, and was inspired by Bruce (<xref ref-type="bibr" rid="B12">1977</xref>). In his description of Central Swedish, lexically contrastive pitch accents occur in the stressed syllable of the word, while a focus-marking H-tone, functionally equivalent to the pitch accents of English, is sequenced after the lexical pitch accent of the last word in the focus constituent. In broad-focus sentences, this focus marking H-tone occurs after the last lexical word and thus before the final boundary L-tone.<xref ref-type="fn" rid="n1">1</xref> Although Pierrehumbert (<xref ref-type="bibr" rid="B53">2000, p. 20</xref>) does not make this equation when she lays out her indebtedness to Bruce (<xref ref-type="bibr" rid="B12">1977</xref>), it is plausible that, despite the difference in functionality of Bruce&#8217;s (<xref ref-type="bibr" rid="B12">1977</xref>) &#8216;sentence accent&#8217; and the &#8216;phrasal accent&#8217; of Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>), the combination of pitch accent and phrase accent was transferred to English nuclear melodies. This ultimately led to boundary tones of an intermediate phrase (L- or H-) in the analysis of Mainstream American English (MAE) known as MAE_ToBI (<xref ref-type="bibr" rid="B6">Beckman &amp; Pierrehumbert, 1986</xref>; <xref ref-type="bibr" rid="B5">Beckman et al., 2005</xref>; <xref ref-type="bibr" rid="B58">Silverman et al., 1992</xref>), whose development from Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) is charted in Ladd (<xref ref-type="bibr" rid="B40">2008, ch. 3</xref>). The second option is to take the entire rise-fall as the pitch accent. An analysis that came close was proposed by &#8217;t Hart, Collier, and Cohen (<xref ref-type="bibr" rid="B62">1990</xref>), whose model used constantly changing line segments as primitives. Their analysis took both the rise and the fall to be pitch accents (&#8216;accent-lending pitch movements&#8217;) and included a convention whereby a syllable may be marked as accented by more than one accent-lending movement, thus making the accent-lending property of any movements that are added to the first in the same syllable vacuous.<xref ref-type="fn" rid="n2">2</xref> Goldsmith (<xref ref-type="bibr" rid="B24">1980</xref>) and Leben (<xref ref-type="bibr" rid="B42">1976</xref>) assumed a tritonal MH*L, whereby the M was deletable in Goldsmith&#8217;s analysis and insertable in Leben&#8217;s, thus taking positions intermediate between the second and third. The third option was unequivocally adopted by Palmer (<xref ref-type="bibr" rid="B49">1922</xref>), who set the stage for the term &#8216;(High) Fall&#8217; for this pitch accent in the British tradition of intonation analysis (<xref ref-type="bibr" rid="B38">Ladd, 1980, ch. 1</xref>). This interpretation was taken over in subsequent descriptions of English intonation as well as in the autosegmental analysis of Gussenhoven (<xref ref-type="bibr" rid="B29">1983</xref>). Apart from the discussion of the position of M by Leben (1997) and Goldsmith (<xref ref-type="bibr" rid="B24">1980</xref>), none of the three options was argued for in a comparison with the other two by any of the authors concerned.</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Three analyses of an accent-lending pitch accent.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74521/"/>
</fig>
<p>Ideally, a phonological account provides a unique transcription for any expression in some language, while there is a unique expression that will correspond to any legitimate transcription. That is, there are no unusable transcriptions (no overanalysis), and there aren&#8217;t any expressions for which no transcription is available (no underanalysis; &#8216;expression&#8217; here refers to an intonation pattern, abstracted away from its morpho-syntactic content). A further property of a successful analysis is its predictive power. Analytical decisions about the tone structure may imply a particular prosodic constituent structure. For instance, a decision to transcribe a steeply falling pitch movement as resulting from a H* followed by a phrasal L-tone in English brings the prediction with it that falling pitch movements are phrase-final.</p>
<p>The purpose of this contribution is to show that examples of underanalysis, overanalysis, and incorrect boundary prediction can be found in MAE_ToBI. Section 2 restates the grammar of MAE_ToBI, including the conventions governing the phonetic implementation (2.1), and considers what that analysis would look like under an off-ramp view (2.2). Section 3 discusses a number of cases of underanalysis in MAE_ToBI, while section 4 does the same for cases of overanalysis. Section 5 identifies a boundary prediction and reports a perception experiment whose results indicate the incorrectness of that prediction. Section 6 reviews empirical evidence presented earlier that bears on the choice between an on-ramp and an off-ramp analysis. Section 7 finally summarizes the off-ramp grammar and discusses the implications of our findings.</p>
</sec>
<sec>
<title>2 MAE_ToBI and its off-ramp alternative</title>
<sec>
<title>2.1. The MAE_ToBI grammar</title>
<p>MAE_ToBI uses four tone paradigms. Addressing them from early to late, there is an optional initial boundary tone of the intonational phrase (IP), five pitch accents to be used for accented syllables, two final boundary tones of the intermediate phrase (ip), and two final boundary tones of the IP. These are listed in (1). In addition, optional downstep applies to any H-tone other than H% (notated !H), provided another H-tone precedes in the IP. In (2), five phonetic implementation conventions applicable to (1) are listed.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(1)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Initial IP-boundary:</p></list-item>
</list>
<list list-type="word">
<list-item><p>%H (optional)</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Pitch accents:</p></list-item>
</list>
<list list-type="word">
<list-item><p>H*, L*, L+H*, L*+H, H+!H*</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>c.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Final ip-boundary:</p></list-item>
</list>
<list list-type="word">
<list-item><p>H-, L-</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>d.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="word">
<list-item><p>Final IP-boundary:</p></list-item>
</list>
<list list-type="word">
<list-item><p>H%, L%</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(2)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>The <italic>F<sub>0</sub></italic> between adjacent targets is obtained by linear interpolation, except for targets of T-, which are &#8216;spread&#8217; between the pitch accent on the left and the boundary on the right.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>H% after H- is upstepped to extra high.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>c.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>L% after H- is upstepped to the value of H-.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>d.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>H*, trailing H, and H- are optionally downstepped relative to a preceding H.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>e.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>The pitch between adjacent H*&#8217;s sags.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>Convention (2c) has a special position. A phonetic implementation rule will not categorically assimilate a tone to another tone, but rather raise or lower a tone&#8217;s target such that its identity is detectable in the signal. However, (2c) leaves no trace of L% and thus effectively turns the phonetic implementation into a mechanism for deleting tones. Ignoring this point, we can calculate the number of two-accent IP-contours by multiplying 2 initial boundary conditions (optional %H) by 5 (prenuclear pitch accents) by 5 (nuclear pitch accents) by 2 (phrase tones) by 2 (IP-tones) = 200 contours. To these, we should add the downstepped contours. If there is no initial %H, 64 contours will have a H in both pitch accents (4 &#215; 4 &#215; 2 &#215; 2), and an additional 16 will have a single pitch accent with H followed by H- (1 &#215; 4 &#215; 2 &#215; 2). When %H is used, only 1 &#215; 1 &#215; 1 &#215; 2 = 2 will <italic>not</italic> have a following H tone. This puts the total number of downstepped contours at 80 + 98 = 178, making for a total of 378 two-accent contours.</p>
</sec>
<sec>
<title>2.2 An off-ramp alternative</title>
<p>Pierrehumbert&#8217;s (<xref ref-type="bibr" rid="B51">1980</xref>) decision to analyze a rising-falling accent-lending contour as a L+H* pitch accent followed by an extraneous L-tone was referred to as an &#8216;on-ramp analysis&#8217; in Gussenhoven (<xref ref-type="bibr" rid="B31">2004: 127</xref>). An &#8216;off-ramp&#8217; analysis will assume a H*+L pitch accent preceded by an extraneous L-tone. A crucial difference between the MAE_ToBI on-ramp analysis and an off-ramp analysis lies in the number of targets that are needed after the nuclear pitch accent. After a pitch peak, two further targets may occur in English, a low target followed by a high target at the IP-end; while after a low valley, there can follow a mid target and high target at the IP-end. To represent these post-peak and post-valley targets, MAE_ToBI provides T- and T%. Most contours, however, have only a single such target. Table <xref ref-type="table" rid="T1">1</xref> lists the eight MAE_ToBI contours with single-tone H* and L* pitch accents in the first eight rows. It shows two post-nuclear targets for contours 2 and 5, the other six having a single overt target after T* (contours 1, 3, 4, 6, 7, and 8). If we simply leave out tones with abstract targets, we produce the representations in column 3.</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>Representations of nuclear contours in MAE_ToBI (column 1) with graphic phonetic implementations, after Pierrehumbert <xref ref-type="bibr" rid="B51 ">1980</xref>(column 2). Column 3 repeats the representations without tones that have no overt target. Column 4 gives representations in an off-ramp analysis without phrase tones and with optional IP-boundary tones.</p>
</caption>
<table>
<tr>
<th align="left"></th>
<th align="left">MAE_ToBI</th>
<th align="left"></th>
<th align="left">MAE_ToBI (overt tones only)</th>
<th align="left">Off-ramp alternative</th>
</tr>
<tr>
<td colspan="5">
<hr/></td>
</tr>
<tr>
<td align="left">1</td>
<td align="left">H* H- H%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74522/"/></td>
<td align="left">H* H%</td>
<td align="left">H* H%</td>
</tr>
<tr>
<td align="left">2</td>
<td align="left">H* L- H%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74523/"/></td>
<td align="left">H* L-H%</td>
<td align="left">H*L H%</td>
</tr>
<tr>
<td align="left">3</td>
<td align="left">H* H- L%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74524/"/></td>
<td align="left">H*</td>
<td align="left">H*</td>
</tr>
<tr>
<td align="left">4</td>
<td align="left">H* L- L%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74525/"/></td>
<td align="left">H* L%</td>
<td align="left">H*L L%</td>
</tr>
<tr>
<td align="left">5</td>
<td align="left">L* H- H%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74526/"/></td>
<td align="left">L* H-H%</td>
<td align="left">L*H H%</td>
</tr>
<tr>
<td align="left">6</td>
<td align="left">L* L- H%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74527/"/></td>
<td align="left">L* H%</td>
<td align="left">L* H%</td>
</tr>
<tr>
<td align="left">7</td>
<td align="left">L* H- L%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74528/"/></td>
<td align="left">L* H-</td>
<td align="left">L*H</td>
</tr>
<tr>
<td align="left">8</td>
<td align="left">L* L- L%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74529/"/></td>
<td align="left">L*</td>
<td align="left">L*</td>
</tr>
<tr>
<td align="left">9</td>
<td align="left"></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74530/"/></td>
<td align="left"></td>
<td align="left">H*L</td>
</tr>
<tr>
<td align="left">10</td>
<td align="left"></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74531/"/></td>
<td align="left"></td>
<td align="left">H* L%</td>
</tr>
<tr>
<td align="left">11</td>
<td align="left"></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74532/"/></td>
<td align="left"></td>
<td align="left">L* L%</td>
</tr>
<tr>
<td align="left">12</td>
<td align="left">L*+H L- L%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74533/"/></td>
<td align="left">L*+H L%</td>
<td align="left">L*H L%</td>
</tr>
</table>
</table-wrap>
<p>Still concentrating on contours 1 to 8, column 5 presents the off-ramp analysis, in which the MAE_ToBI ip-boundary tones have been absorbed as trailing tones in the pitch accent, except for contour 4, which has a trailing L in column 5, but no corresponding L- in column 4.</p>
<p>These off-ramp versions amount to a system with four pitch accents (H*, H*L, L*, L*H) and an IP-optional boundary tone. Spelling out the 12 representations by combining these four pitch accents with the three boundary conditions H%, L%, and &#216; (no tone) yields four further contours. The representation H*L L% for contour 4 contrasts with contour 9, the &#8216;half-completed fall&#8217;, and contour 10, the &#8216;High Level-Slump&#8217;, and 12, the delayed fall, to which we turn in Section 4.4. A discussion of contour 11 appears in Section 3.2.</p>
</sec>
</sec>
<sec>
<title>3 Underanalysis in MAE_ToBI</title>
<sec>
<title>3.1 Contours ending in mid pitch</title>
<p>Pierrehumbert (<xref ref-type="bibr" rid="B51">1980, p. 88</xref>) discussed the contrast between mid-ending (3) (cf. Pierrehumbert&#8217;s Figure 6.4) and (4), noting that her analysis had a single representation for them. She argues that contour (4) is a &#8216;chanted&#8217; version of (3) and that chanted speech is an orthogonal variable not requiring a separate tonal representation. MAE_ToBI notates them as H* !H- L%. Against this view, Hayes and Lahiri (<xref ref-type="bibr" rid="B35">1991</xref>) showed that the English vocative chant requires a representation which accounts for the neutralization of vowel quantity contrast in IP-final syllables, causing <italic>Je-en!</italic> and <italic>Ja-ane</italic>! to be prosodically identical. Moreover, !H- crucially requires a syllabic association to a post-accentual stressed syllable, since its phonetic alignment is with -<italic>nath</italic>- in (3) rather than with either the preceding or following unstressed syllable (<xref ref-type="bibr" rid="B37">Ladd, 1978</xref>; <xref ref-type="bibr" rid="B43">Liberman, 1975</xref>). Example (3) could be a tentative suggestion (<xref ref-type="bibr" rid="B19">Crystal, 1969, p. 147</xref>; <xref ref-type="bibr" rid="B23">Gibbon, 1976, p. 135</xref>; <xref ref-type="bibr" rid="B29">Gussenhoven, 1983, p. 40</xref>; <xref ref-type="bibr" rid="B65">Uldall, 1961</xref>) or be used to chide someone. These effects are quite different from that of (4). That is, the contrast between (3) and (4) represents a genuine case of underanalysis in MAE_ToBI.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(3)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74534/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx3">http://dx.doi.org/10.5334/labphon.30.wavEx3</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(4)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74535/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx4">http://dx.doi.org/10.5334/labphon.30.wavEx4</ext-link></p>
<p>The off-ramp analysis provides H*L &#216;, contour 9, for (3), which contrasts with the rapid final fall, contour (4). By assuming that trailing L has mid-low pitch, while L% is pronounced at fully low pitch, the two L-tones in contour 4 acquire overt tonal targets. Also, the mid-low ending of (3) is explained by the pronunciation of trailing L at a point near the IP-boundary.</p>
<p>Vocative chants require an additional pitch accent, notated H*+H in Gussenhoven (<xref ref-type="bibr" rid="B31">2004, ch. 15</xref>). It is given in Table <xref ref-type="table" rid="T2">2</xref>, as H*H, where also the falling-rising vocative chant (<xref ref-type="bibr" rid="B29">Gussenhoven, 1983, p. 41</xref>, with reference to <xref ref-type="bibr" rid="B51">Pierrehumbert, 1980</xref>) and the low falling vocative chant (<xref ref-type="bibr" rid="B31">Gussenhoven, 2004, p. 315</xref>) are included, and accounted for by the addition of H% and L%, respectively. After the extension of the off-ramp grammar with this H*H pitch accent, contour 13 takes care of (4). In addition, we generate representations for two further vocative chants.</p>
<table-wrap id="T2">
<label>Table 2</label>
<caption>
<p>Representation of the vocative chant in MAE_ToBI (column 2) with graphic phonetic implementations after Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) and representations for the mid-falling, falling-rising, and low vocative chants in the off-ramp analysis.</p>
</caption>
<table>
<tr>
<th align="left"></th>
<th align="left">MAE_ToBI</th>
<th align="left"></th>
<th align="left">MAE_ToBI (overt tones only)</th>
<th align="left">Off-ramp alternative</th>
</tr>
<tr>
<td colspan="5">
<hr/></td>
</tr>
<tr>
<td align="left">13</td>
<td align="left">H* !H- L%</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74536/"/></td>
<td align="left">H* !H-</td>
<td align="left">H*H</td>
</tr>
<tr>
<td align="left">14</td>
<td align="left"></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74537/"/></td>
<td align="left"></td>
<td align="left">H*H H%</td>
</tr>
<tr>
<td align="left">15</td>
<td align="left"></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74538/"/></td>
<td align="left"></td>
<td align="left">H*H L%</td>
</tr>
</table>
</table-wrap>
<p>There is in fact a third mid-ending contour, for which MAE_ToBI would equally have to use H* !H-L%. Contour 10 is part of class of contours ending in a fall to mid after a high stretch beginning after the accented syllable. We will return to these contours in Section 4.4.</p>
</sec>
<sec>
<title>3.2 Scathing intonation</title>
<p>The second case of a missed contrast concerns contour 8, L* L- L%, the &#8216;scathing&#8217; contour, as it was called by Alex Monaghan in a now defunct Linguist List message. It is an echo-statement, typically used as a repetition of a listener&#8217;s earlier utterance, used to express disparagement and disbelief. Gussenhoven (<xref ref-type="bibr" rid="B31">2004, p. 301</xref>) claimed that there are two &#8216;scathing intonations&#8217;. One remains level from the low-pitch accented syllable onwards, which has the force of <italic>Here we go again!</italic>, a &#8216;routine&#8217; meaning identified by Ladd (<xref ref-type="bibr" rid="B37">1978</xref>), shown in panel (a) of Figure <xref ref-type="fig" rid="F2">2</xref>.<xref ref-type="fn" rid="n3">3</xref> The other contour descends somewhat within a low register. It may express a stronger degree of mockery, as in panel (b), but has other uses too, as in Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>, Figure 4.19), where it is used on <italic>damn</italic> after H* on <italic>God</italic> in <italic>God damn it!</italic> The off-ramp analysis transcribes these as L* &#216; and L* L%, respectively.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Two &#8216;scathing&#8217; contours, a low level contour on <italic>It&#8217;s your MOTHers fault again</italic> (panel a, male GBE speaker) and the low falling contour on <italic>WHO broke the dish!</italic> (panel b, speaker CG). From Gussenhoven (<xref ref-type="bibr" rid="B31">2004</xref>). This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav2a">http://dx.doi.org/10.5334/labphon.30.wav2a</ext-link> and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav2b">http://dx.doi.org/10.5334/labphon.30.wav2b</ext-link></p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74539/"/>
</fig>
</sec>
<sec>
<title>3.3 H+!H*, but no H+L*</title>
<p>The third case of a missed contrast was pointed out to me by Bruce Hayes with reference to Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) (personal communication, July 1991) and concerns the contrast between downstepped !H* and L* after a preceding high syllable. MAE_ToBI provides H+!H* to cover the first case, but since there is no H+L*, it cannot describe the second.<xref ref-type="fn" rid="n4">4</xref> Grice (<xref ref-type="bibr" rid="B26">1995</xref>) independently treated this distinction, exemplified by her with (5) and (6), pointing out that these contours required the adoption of a generally applicable leading H, which is prefixed to either L* or H*. In (5), the accented syllable -<italic>ma</italic>- is fully low-pitched, due to L*, while that in (6) is mid-pitched, as for a downstepped !H*. Illustrative contours are presented in Figure <xref ref-type="fig" rid="F3">3</xref>. Possibly, the slowly rising pitch towards H% in the contour in panel (a) serves an enhancement of L*. The contrast was included in the analysis of German by Grice, Baumann, and Benzm&#252;ller (<xref ref-type="bibr" rid="B27">2005</xref>).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(5)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74540/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx5">http://dx.doi.org/10.5334/labphon.30.wavEx5</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(6)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74541/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx6">http://dx.doi.org/10.5334/labphon.30.wavEx6</ext-link></p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>A L* target (panel a, female GBE speaker) and a downstepped !H* target (panel b, female Mid-Western MAE speaker) with leading H&#8217;s on <italic>tomatoes</italic> in <italic>The tomatoes haven&#8217;t arrived yet</italic>. The contour in panel (a) is an echo question, that in panel (b) can be used as a statement. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav3a">http://dx.doi.org/10.5334/labphon.30.wav3a</ext-link> and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav3b">http://dx.doi.org/10.5334/labphon.30.wav3b</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74542/"/>
</fig>
<p>Following Grice (<xref ref-type="bibr" rid="B26">1995</xref>), the off-ramp analysis assumes a prefix H, here notated in italic font to separate it from the base pitch accent (<italic>H</italic>H*L, <italic>H</italic>L*, etc.). Strikingly, H* is invariably downstepped after the pre-accentual peak (<xref ref-type="bibr" rid="B26">Grice, 1995, p. 202</xref>). The generalization that arises from the pronunciation of this pitch accent and of the vocative chant (Section 3.1) is that within pitch accents downstep is obligatory. Under an assumption of &#8216;P(itch) A(ccent)-internal downstep&#8217; (<xref ref-type="bibr" rid="B31">Gussenhoven, 2004, p. 301</xref>), the inclusion of H* and its leading H in the same pitch accent renders the downstep inevitable, quite as in the case of H*H, the vocative chant. When trigger H and target H* are not contained within a pitch accent, downstep is optional, but while any H-tone can be the trigger, only H* can be downstepped. Against the background of these generalizations in the off-ramp analysis, the postulation of downstep of the phrasal tone in the chanted call in MAE_ToBI now looks arbitrary, since in other contexts no contrastive downstep on H- is in evidence. For instance, there has been no demonstration that H* L- H% (high-low-high) is categorically distinct from H* !H- H% (high-mid-high).</p>
</sec>
<sec>
<title>3.4 Virtual vs. real leading H</title>
<p>My fourth case has not been discussed before, as far as I am aware. To describe high level pitch between a high and a downstepped high pitch accent, ToBI uses a prenuclear H* which is followed by H+!H*, where the high stretch between the pitch accents is described as an interpolation between H* and leading H. An example of this contour (cf. &#8217;t Hart, Collier, &amp; Cohen&#8217;s [<xref ref-type="bibr" rid="B62">1990</xref>] &#8216;flat hat&#8217;) is shown in Figure <xref ref-type="fig" rid="F4">4</xref>, panel (a). The general descending profile is a common, though not a necessary feature of this contour. The MAE_ToBI analysis implies that there is no transcription available for the same contour with an upstepped high pitch on the syllable before the second accented syllable, as in the contour in panel (b). This contrast seems quite categorical, with a distinct note of liveliness in contour (b) which is absent in contour (a).</p>
<fig id="F4">
<label>Figure 4</label>
<caption>
<p>A descending &#8216;flat hat&#8217; contour without (panel a) and a &#8216;flat hat&#8217; contour with a raised peak on the syllable before the second accented syllable (panel b). Male Canadian English speaker. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav4a">http://dx.doi.org/10.5334/labphon.30.wav4a</ext-link> and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav4b">http://dx.doi.org/10.5334/labphon.30.wav4b</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74543/"/>
</fig>
<p>To account for the difference between the contours in Figure <xref ref-type="fig" rid="F4">4</xref>, we must assume that the pronunciation of the target of prenuclear H* continues until just before the first tone in the nuclear pitch accent. In fact, contours 3 and 8 already made it clear that tonal targets are continued rightwards if there is no further tonal target in the IP: without any following tones, a string-final H* is realized as high level pitch until end of the IP, while string-final L* in the same position produces low level pitch. Similarly, a trailing tone is continued when string-final, as in contours 7 and 13. This &#8216;continuation&#8217; of tones appears to apply generally to any English morpheme-final tone. In addition to the situation before a toneless boundary, there are three inter-morphemic stretches in which this continuation occurs:</p>
<list list-type="roman-lower">
<list-item>
<p>from a boundary tone to a pitch accent;</p>
</list-item>
<list-item>
<p>between pitch accents;</p>
</list-item>
<list-item>
<p>from a pitch accent to a boundary tone.</p>
</list-item>
</list>
<p>MAE_ToBI presents the continuation of tonal targets as an anomaly, applicable only to the phrase tone, i.e., the equivalent of context (iii). The most widely discussed case here is that of L- between H* and H%, which forms a &#8216;floodplain&#8217;, in the terminology of Lickley et al. (<xref ref-type="bibr" rid="B44">2005</xref>), but the same is true for mid-level stretches in the MAE_ToBI L* H- H% contour. From the off-ramp perspective, these anomalies disappear as part of the generalization that unspecified inter-morphemic stretches are filled with the tone on the left. Thus, prenuclear H* in (7) continues its pronunciation from -<italic>ron-</italic> onwards, until preparations need to be made for the pronunciation of the downstepped target of !H*. To account for this continued pronunciation, Gussenhoven (<xref ref-type="bibr" rid="B30">2000</xref>) introduced the concept of double alignment. Alignment with other phonological constituents quite generally determines the location of tonal targets (cf. <xref ref-type="bibr" rid="B46">McCarthy &amp; Prince, 1993</xref>). It is expressed as a coincidence of the edges of two constituents, such as when a prefix is said to align its left edge with the left edge of the word it attaches to. Thus, an initial boundary tone aligns its left edge with the left edge of the IP, a final boundary tone aligns its right edge with the right edge of the IP, a leading H aligns its right edge with the left edge of the following T*, and an associated tone aligns with an edge of the accented syllable rime (cf. <xref ref-type="bibr" rid="B52">Pierrehumbert, 1993</xref>). Unspecified space between targets is covered by an interpolation between them in MAE_ToBI, following Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>). Double alignment means that the left-hand target additionally acquires a right-hand target, since the tone is both left-aligned and right-aligned, the latter being shown as empty bullets (<xref ref-type="bibr" rid="B68">van de Ven &amp; Gussenhoven, 2011</xref>). In (8), leading H now defines a contour distinct from (7), one with raised pitch on the syllable immediately before the nuclear accent. A contour like (8) is reported for <italic>Now you&#8217;re CURving to the RIGHT</italic> in Figure <xref ref-type="fig" rid="F1">1</xref> in Shattuck-Hufnagel et al. (<xref ref-type="bibr" rid="B57">2004</xref>), where I interpret the mid target on <italic>CUR-</italic> to be a realization of H* and <italic>the</italic> to be the location of leading H. In their small corpus, 39% of two-peak contours had an intervening peak on an unstressed syllable, many of which are likely to be further examples.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(7)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74544/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx7">http://dx.doi.org/10.5334/labphon.30.wavEx7</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(8)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74545/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx8">http://dx.doi.org/10.5334/labphon.30.wavEx8</ext-link></p>
</sec>
<sec>
<title>3.5 Prefix L* and L*<italic>+</italic>H</title>
<p>Contour 12 in Table <xref ref-type="table" rid="T1">1</xref> raises two issues in the intonational phonology of English, corresponding to two contour classes which have a low-pitched accented syllable followed by a rising-falling contour, <italic>viz.</italic> &#8216;delayed&#8217; contours and contours ending in a &#8216;slump&#8217;. The MAE_ToBI representation belongs to the first class. It was characterized as having &#8216;scoop&#8217; by Vanderslice &amp; Pierson (<xref ref-type="bibr" rid="B67">1967</xref>) with reference to Hawaiian English. For American English, Vanderslice (<xref ref-type="bibr" rid="B66">1972, p. 1053</xref>) notes that scoop, which corresponds to Ladd&#8217;s &#8216;scooped&#8217; or &#8216;delayed peak&#8217; contours (<xref ref-type="bibr" rid="B38">Ladd, 1980</xref>, <xref ref-type="bibr" rid="B40">2008</xref>) and my own [Delay] (<xref ref-type="bibr" rid="B29">Gussenhoven, 1983</xref>), &#8216;delays the upward pitch obtrusion associated with an accented syllable&#8217;. Semantically, it has been characterized as having an intensifying (<xref ref-type="bibr" rid="B48">O&#8217;Connor &amp; Arnold, 1973, p. 78</xref>; <xref ref-type="bibr" rid="B63">Tench, 1996, p. 126</xref>, among others) or dominating effect (<xref ref-type="bibr" rid="B11">Brazil, 1985, p. 129</xref>), or as expressing that the speaker is impressed (<xref ref-type="bibr" rid="B69">Wells, 2006, pp. 218, 221</xref>). These scooped or delayed contours can be captured by a prefix L*-tone, to be inserted to the left of H* (<xref ref-type="bibr" rid="B31">Gussenhoven, 2004, p. 307</xref>). Prefixal <italic>L*</italic>, notated in italic font, associates with the accented syllable, dislodging following H*, whose asterisk is now left out. It may combine with prefix-H. Following our discussion of Obligatory PA-internal downstep, the presence of leading H in the pitch accent implies downstep on the <italic>F<sub>0</sub></italic> peak due to underlying H*, located on <italic>market</italic> in (9). That is, no contrast between a downstepped and non-downstepped second peak in (9) is expected (<xref ref-type="bibr" rid="B31">Gussenhoven, 2004, p. 321</xref>). In (9), there is low pitch on <italic>To</italic>, high pitch on <italic>the</italic>, late rising pitch on <italic>mar-</italic>, and falling pitch on <italic>-ket</italic>.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(9)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74546/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx9">http://dx.doi.org/10.5334/labphon.30.wavEx9</ext-link></p>
<p>The question arises then whether the existence of the simplex pitch accent L*H in the off-ramp analysis by the side of a prefix-L* attaching to H* represents a case of overanalysis, i.e., whether L*H is equivalent to <italic>L*</italic>H. There are two arguments for considering them to be contrasting representations. In <italic>L*</italic>H, H has the status of a dislodged H*-tone, which retains the properties of H*. This means, first, that it is not treated as the last tone of a pitch accent, which would require it to align with the next pitch accent, like H in monomorphemic L*H, but rather will continue its pronunciation until the next pitch accent, creating high level pitch. Second, downstep targets H*-tones, predicting that L*-prefixed H-tones (i.e., underlying H*-tones), but not trailing H-tones, can be contrastively downstepped. So while <italic>L*</italic>H has a counterpart <italic>L*</italic>!H, there should be no L*!H. Example (10) illustrates a prenuclear <italic>L*</italic>!H, in which the H-tone creates mid level pitch, before two occurrences of <italic>L*</italic>!HL. This contour is predicted to contrast with a non-downstepped version. In contradistinction to (10), contour (11) has two occurrences of L*H in prenuclear position, predicting that the pitch between <italic>back</italic> and <italic>boy</italic> is a slow rise, and also that the pitch on <italic>-sty</italic> of <italic>nasty</italic> is not contrastively mid or high.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(10)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74547/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx10">http://dx.doi.org/10.5334/labphon.30.wavEx10</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(11)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74548/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx11">http://dx.doi.org/10.5334/labphon.30.wavEx11</ext-link></p>
<p>An empirical argument may be based on meaning. An eye-tracking study in fact suggested that <italic>L*</italic>HL is associated with newness, just like H*L, while L*H is associated with givenness (<xref ref-type="bibr" rid="B14">Chen et al., 2007</xref>). Yet, the above claims evidently require more empirical research before it can be decided whether the off-ramp analysis is here running into a case of overanalysis or whether we here have another case of underanalysis in MAE_ToBI.</p>
<p>The second class of rise-fall contours end in a &#8216;slump&#8217;, a truncated type of final fall, which is characteristic of Northern British English contours, variants of which are surveyed in Cruttenden (<xref ref-type="bibr" rid="B18">1997, ch. 5</xref>). Nolan and Grabe (<xref ref-type="bibr" rid="B47">1997</xref>) pointed out that Pierrehumbert&#8217;s convention of using H- L% to mean mid pitch (MAE_ToBI&#8217;s !H-L%, see convention (2c)) makes it impossible to use H- L% to describe the slump. This type of contour, to be sure, has not been included in descriptions of MAE, and MAE_ToBI cannot be criticized for failing to provide a representation for the Rise-Level-Slump of Northern Irish English on which Nolan and Grabe (<xref ref-type="bibr" rid="B47">1997</xref>) base their case. Indeed, Mayo et al. (<xref ref-type="bibr" rid="B45">1997</xref>) abandon clause (2c) so as to use H-L% for the &#8216;slump&#8217; in Glaswegian English. Yet, it may be argued that a phonology of a complex intonation system like that for English may generate contours that do not occur in all varieties. As noted by Pierrehumbert (<xref ref-type="bibr" rid="B53">2000, p. 27</xref>), &#8216;nothing like the full set generated by [MAE_ToBI] has ever been documented&#8217;. A large proportion of a grammar&#8217;s legitimate contours may never be encountered in anyone&#8217;s lifetime, any more than will be the majority of morphosyntactic structures generated by some simple mini-grammar of English. Such non-occurrence may well be interpreted as absence from the grammar, provided it takes the form of a stochastic algorithm (<xref ref-type="bibr" rid="B20">Dainora, 2006</xref>). Either way, varieties are likely to differ in the frequency with which certain structures are used for certain pragmatic functions (<xref ref-type="bibr" rid="B25">Grabe &amp; Post, 2004</xref>; <xref ref-type="bibr" rid="B56">Ritchart &amp; Arvaniti, 2014</xref>), while there will also be cases of absolute non-use (see also <xref ref-type="bibr" rid="B16">Cole &amp; Shattuck-Hufnagel, 2016</xref>). Wells (<xref ref-type="bibr" rid="B69">2006, p. 245 fn 8</xref>), for instance, notes that the second edition of O&#8217;Connor and Arnold (<xref ref-type="bibr" rid="B48">1973</xref>) was the first British English course book that awarded the Fall-rise (H*L H%) full treatment as a neutral polar question contour. Earlier, it had not been reported for questions in GBE, and in MAE it is apparently (still?) not used in that function. Or again, I have found it hard to elicit H* H% contours from GBE speakers, who tend to produce L*H H% instead, and the speaker of the contour in Figure <xref ref-type="fig" rid="F3">3</xref>, panel (b), associated it with GBE, while having no problem producing it. Be this as it may, the off-ramp analysis readily provides representations for slumped contours by providing L% after pitch accents like H* and L*H, as shown in (12a), which contrasts with (12b) of the standard languages. The off-ramp analysis offers contour 10 in Table <xref ref-type="table" rid="T1">1</xref> for a high-beginning equivalent of (12a), a contour which has not been reported even for northern British English. Arguably, therefore, we are here dealing with a systematic case of overgeneration. However, there is a difference between this case and the cases of overanalysis in MAE_ToBI to be discussed in Section 4. The MAE_ToBI cases concern putative contrasts that one would not expect to turn up in any variety of English, while contour 10 in Table <xref ref-type="table" rid="T1">1</xref>, being clearly distinct from other contours, might.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(12)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74549/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx12a">http://dx.doi.org/10.5334/labphon.30.wavEx12a</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74550/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx12b">http://dx.doi.org/10.5334/labphon.30.wavEx12b</ext-link></p>
</sec>
</sec>
<sec>
<title>4 Overanalysis in MAE_ToBI</title>
</sec>
<sec>
<title>4.1 Prenuclear L*<italic>+</italic>H</title>
<p>Overanalysis may arise from sequences of H-tones, one or both of which are unstarred. Some of these are given in (13), where the transcriptions to the right of the arrow would not appear to describe a different contour from that on the left.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(13)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Ambiguity of analysis I</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>L*+H</p></list-item>
</list>
<list list-type="word">
<list-item><p>H+!H*</p></list-item>
</list>
<list list-type="word">
<list-item><p>&#8660;</p></list-item>
</list>
<list list-type="word">
<list-item><p>L* H+!H*</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>L*+H</p></list-item>
</list>
<list list-type="word">
<list-item><p>H*</p></list-item>
</list>
<list list-type="word">
<list-item><p>&#8660;</p></list-item>
</list>
<list list-type="word">
<list-item><p>L* &#160; H*</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>c.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>L*+H</p></list-item>
</list>
<list list-type="word">
<list-item><p>H- H%</p></list-item>
</list>
<list list-type="word">
<list-item><p>&#8660;</p></list-item>
</list>
<list list-type="word">
<list-item><p>L* H- H%</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>Figure <xref ref-type="fig" rid="F5">5</xref> presents <italic>F<sub>0</sub></italic> contours on <italic>toRONto is the capital of onTArio</italic> for cases (13a) in panels (a) and (b) and for cases (13b) in panels (c) and (d). The contours in panels (a) and (c) might at first sight be transcribed as on the left of the arrow, while those in panels (b) and (d), in which the mid sections have been resynthesized, might be expected to be transcribed with the symbols to the right of the arrow. However, the original and the resynthesized contours are not easily interpretable as representing different intonations. Chapter 5 in Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) discusses these complications of the analysis and characterizes the contours in panels (b) and (d) as &#8216;impossible&#8217;. Like case (11c), these ambiguities are an inevitable consequence of her analysis.</p>
<fig id="F5">
<label>Figure 5</label>
<caption>
<p>A pre-nuclear L*-beginning rise in <italic>But ToRONto is the capital of onTArio</italic> and a resynthesized version with an accelerated early part of the rising movement before H+!H* (panels a, b) and before (non-downstepped) H* (panels c, d). Female GBE speaker. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav5a">http://dx.doi.org/10.5334/labphon.30.wav5a</ext-link>, <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav5b">http://dx.doi.org/10.5334/labphon.30.wav5b</ext-link>, <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav5c">http://dx.doi.org/10.5334/labphon.30.wav5c</ext-link>, and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav5d">http://dx.doi.org/10.5334/labphon.30.wav5d</ext-link></p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74551/"/>
</fig>
<p>We begin by observing that the absence of a L+H* pitch accent in the off-ramp analysis forces it to interpret inter-accentual slow falls as instances of H*L and inter-accentual slow rises as instances of L*H. This is shown graphically in (14) for prenuclear H*L. The low target before the nuclear <italic>F<sub>0</sub></italic> peak is described by aligning the trailing L rightwards, thus moving its target to a point just before the target of the next tone. The space between the targets of H* and L is filled by an interpolation. In MAE_ToBI, which lacks H*+L but has L+H*, the slow fall is an interpolation between a prenuclear H* and a leading L of the next pitch accent.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(14)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74552/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The off-ramp view thus suggests two things. One is that trailing tones of <italic>prenuclear</italic> pitch accents are aligned rightmost, i.e., with the left edge of the first tone of the next pitch accent. The trailing L of the <italic>nuclear</italic> pitch accent is aligned leftmost, i.e., defines a rapid fall. The second implication is that linear interpolations are restricted to tones <italic>within the same pitch accent</italic>. This intra-morphemic linear interpolation between tones thus contrasts with the inter-morphemic continuation of tonal targets over stretches of speech between pitch accents and boundary tones (see Section 3.4). By having stretchable interpolations between tones in pre-nuclear pitch accents, we guarantee that H* and H*L will be distinct in all positions and all contexts, as will L* and L*H.<xref ref-type="fn" rid="n5">5</xref> The slow rises in panels (a) and (b) of Figure <xref ref-type="fig" rid="F5">5</xref>, therefore, are described by a pre-nuclear L*H followed by <italic>H</italic>H*, whereby contour (b) is a less felicitous realization of that representation. Likewise, those in panels (c) and (d) are described by L*H before H*. The overgeneration in (11c) similarly disappears, both of them corresponding to L*H H%.</p>
<p>Like MAE_ToBI, which has an optional %H, the off-ramp analysis offers two choices at the initial boundary of the IP, notated as %L for low-to-mid pitch and %H for high pitch. The contours in Figure <xref ref-type="fig" rid="F5">5</xref> have %H. Figure <xref ref-type="fig" rid="F6">6</xref> shows rising prenuclear stretches after %L. Since MAE_ToBI uses leading H in H+!H* merely to provide a high target in the syllable before !H*, as we saw in Section 4.1 above, it is unclear which of the four available transcriptions (L* H*, L*+H H*, L* H+!H*, or L+H H+!H*) describes which contour in Figure <xref ref-type="fig" rid="F6">6</xref>.</p>
<fig id="F6">
<label>Figure 6</label>
<caption>
<p>Two slow rises from the pre-nuclear accent in <italic>But the SECond of these is the BEST</italic> to a nuclear !H* without leading H (panel a) and <italic>H</italic>H* (i.e., with leading H, panel b). Female speaker of GBE. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav6a">http://dx.doi.org/10.5334/labphon.30.wav6a</ext-link> and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav6b">http://dx.doi.org/10.5334/labphon.30.wav6b</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74553/"/>
</fig>
<sec>
<title>4.2 L*<italic>+</italic>H vs. H*</title>
<p>In (15), a source of ambiguity is given which has been widely commented on (<xref ref-type="bibr" rid="B1">Arvaniti, 2016</xref>; <xref ref-type="bibr" rid="B31">Gussenhoven, 2004, p. 319</xref>; <xref ref-type="bibr" rid="B40">Ladd, 2008, p. 96, fn 3</xref>; <xref ref-type="bibr" rid="B54">Pitrelli et al., 1994</xref>), in particular for occurrences on the first syllable of the IP, case (15a). In (15), the accolade represents the initial IP boundary.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(15)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Ambiguity of analysis II</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>{ &#8230; L+H*</p></list-item>
</list>
<list list-type="word">
<list-item><p>&#8660;</p></list-item>
</list>
<list list-type="word">
<list-item><p>{ &#8230; H*</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>{ H* &#8230; L+H*</p></list-item>
</list>
<list list-type="word">
<list-item><p>&#8660;</p></list-item>
</list>
<list list-type="word">
<list-item><p>{ H*&#8230; H* (assuming &#8216;sagging&#8217;)</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>c.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="word">
<list-item><p>{ L+H*</p></list-item>
</list>
<list list-type="word">
<list-item><p>&#8660;</p></list-item>
</list>
<list list-type="word">
<list-item><p>{ H*</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>In cases (15a, b), an unaccented syllable precedes a H*-bearing syllable with leading L. For (15a), the issue is whether the initial unaccented syllables are low enough to warrant the choice of the leading L, in a situation in which low or mid pitch is predicted anyway, given the absence of an initial %H. If the prediction for L+H* is that of a later peak, as suggested by Steedman (<xref ref-type="bibr" rid="B59">1991</xref>), the analysis would imply a three-way peak timing contrast: early peak (H*), later peak (L+H*), and very late or delayed peak (L*+H). This is not, however, a claim that has been explicitly made, as far as I know. Steedman (<xref ref-type="bibr" rid="B59">1991, p. 273</xref>) consistently uses H* in combination with L-L% and L+H* in combination with L- H%, noting that the H* is somewhat later in the second type of contour than in the first, thus interpreting the difference as allophonic. The examples in the MAE_ToBI manual (<xref ref-type="bibr" rid="B4">Beckman &amp; Ayers, 1994</xref>) suggest that for L+H* there is low pitch preceding the peak <italic>in addition to</italic> a high pitched peak, resulting in a wider rising flank than for plain H*. In panels (a) and (b) of Figure <xref ref-type="fig" rid="F7">7</xref>, realizations of <italic>MariANNa won it</italic> are given (from <xref ref-type="bibr" rid="B4">Beckman &amp; Ayers, 1994</xref>; creak occurs on [n&#618;t]). Under this interpretation, the question arises how pronunciations of L* H-H% are to be transcribed that similarly differ in pitch range. The pair of contours in panels (c) and (d) are arguably analyzable as %H L* H-H%, for the narrower range one in panel (c), and as %H L*+H H- H%, for the wider-range pronunciation in panel (d), thus providing a use for the putative contrast in (15a). This approach would however leave yet further pitch range differences unaccounted for, like two heights for the beginning of the fall in H+!H* L-L%. Understandably, no such more general representation of pitch range variation has been included in MAE_ToBI, which views pitch range differences other than downstep as orthogonal to the symbolic transcription (cf. <xref ref-type="bibr" rid="B9">Bolinger, 1951</xref>; <xref ref-type="bibr" rid="B40">Ladd, 2008, p. 36, sec. 5.2</xref>). The putative contrast in (15a), therefore, is an anomalous feature of the analysis, if the interpretation is in terms of pitch range.</p>
<fig id="F7">
<label>Figure 7</label>
<caption>
<p><italic>Marianna won it</italic> with H* L-L% (panel a) and L+H* L-L% (panel b) (female MAE speaker, from <xref ref-type="bibr" rid="B4">Beckman &amp; Ayers, 1994</xref>) and <italic>Manianna won it?</italic> with neutral (panel c) and wide pitch range (panel d) pronunciations of %H L* H- H%. Female GBE speaker. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav7a">http://dx.doi.org/10.5334/labphon.30.wav7a</ext-link>, <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav7b">http://dx.doi.org/10.5334/labphon.30.wav7b</ext-link>, <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav7c">http://dx.doi.org/10.5334/labphon.30.wav7c</ext-link>, and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav7d">http://dx.doi.org/10.5334/labphon.30.wav7d</ext-link></p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74554/"/>
</fig>
<p>The evaluation of case (15b) depends on the assumptions made for the shape of the interpolation between H*-tones in the contour to the left of the arrow and the realization of the leading L-tone. For the first aspect, Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) argued for a sagging transition, instead of a level interpolation, as would be predicted by double alignment assumed in the off-ramp analysis. The second aspect concerns realization of the leading L of L+H* in the left-hand transcription. If sagging is assumed and the realization of L in L+H* is low-pitched, the prediction is that contour (c) in Figure <xref ref-type="fig" rid="F8">8</xref> is a realization of H* L+H*, while contour (b) is the realization of H* H*. The two contours do not, however, appear to represent different intonations, while both are distinct from contour (a).</p>
<fig id="F8">
<label>Figure 8</label>
<caption>
<p><italic>ToRONto is the capital of onTArio</italic> with MAE_ToBI H* &#8230; H* and level pitch (top), H* &#8230; H* and sagging pitch (middle), and H* &#8230; L+H* (bottom), resynthesized versions from a source utterance by a female GBE speaker. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav8a">http://dx.doi.org/10.5334/labphon.30.wav8a</ext-link>, <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav8b">http://dx.doi.org/10.5334/labphon.30.wav8b</ext-link>, and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav8c">http://dx.doi.org/10.5334/labphon.30.wav8c</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74555/"/>
</fig>
<p>The need for a deep valley for the nuclear L+H* after a prenuclear H* was questioned by Ladd and Schepman (<xref ref-type="bibr" rid="B41">2003</xref>), who showed that its depth varied with the distance between the H*-tones. While arguing on the basis of the location of the low target that it in fact derives from a L-tone, they point out that there is no contrast between different depths and thus no contrast between contours (b) and (c), and thus no contrast between L+H* and H* under an assumption of sagging. More realistically, targets of leading tones are implemented by gradient realization rules creating undershoot in Grice (<xref ref-type="bibr" rid="B26">1995, pp. 226&#8211;228</xref>), in the spirit of Chen &amp; Xu&#8217;s (<xref ref-type="bibr" rid="B15">2006</xref>) weak targets, whose realization has less priority than a target of T*, say. Abandoning the requirement of a low realization of leading L as well as the convention of sagging interpolations would enable MAE_ToBI to correctly describe the difference between contours (b) and (c) on the one hand and contour (a) on the other. Without L+H* (or LH*, as I would have notated it), the off-ramp analysis cannot run into this ambiguity between H* H* and H* L+H*. If the pitch is high level, we have a case of H* H*, while a slow fall is described by H*L H*.</p>
<p>Neither can there be any ambiguity between L+H* and H* in the case of an IP-initial accented syllable in the off-ramp analysis (15c). With only H* available as a transcription (abstracting away from the option of a trailing L and downstep on H*) and with %L and %H as initial boundary tones, there are two ways in which preceding pitch may contrast, mid/low pitched (%L) or high pitched (H%), phonetically realized in the onset and early section of the rime. Compare this with the four transcriptions that are available in MAE_ToBI, H*, L+H*, %H H*, and %H L+H*. If unaccented syllables occur before the pitch accent (15a), we may include a leading H before nuclear T* in the off-ramp analysis and before H* in MAE_ToBI. Four transcriptions are now produced in the off-ramp analysis and six in MAE_ToBI. Off-ramp leading H is pronounced higher than a preceding H, including %H, while following H* is obligatorily downstepped (see Section 3.2). For the first three syllables of <italic>The tomatoes</italic>, the four off-ramp options are therefore %L H* or low-low-high, %L <italic>H</italic>H* or low, extra high, downstepped high, %H H* or high, high, high, and %H <italic>H</italic>H* or high, extra high, and downstepped high. MAE_ToBI&#8217;s six patterns have not explicitly been described.</p>
</sec>
</sec>
<sec>
<title>5 An incorrect boundary prediction</title>
<sec>
<title>5.1 Boundaries in MAE_ToBI</title>
<p>MAE_ToBI specifies two prosodic boundaries. First, a T- without a following T% indicates the end of an intermediate phrase (ip), while any T-T% combination additionally indicates the end of an intonational phrase (IP).<xref ref-type="fn" rid="n6">6</xref> As observed by Ladd (<xref ref-type="bibr" rid="B40">2008, p. 107</xref>), MAE_ToBI would appear to predict boundaries where there are none. As a result of the absence of any falling (H*+L or H+L*) pitch accents, a sharpish accent-lending fall minimally predicts an ip-boundary, since that fall can only be described by H* followed by an ip-boundary tone, L-. The incorrectness of this implication is suggested by a contour type that is not often discussed in the literature on English (but see <xref ref-type="bibr" rid="B18">Cruttenden, 1997, pp. 59, 76</xref>; <xref ref-type="bibr" rid="B29">Gussenhoven, 1983, p. 35</xref>; <xref ref-type="bibr" rid="B40">Ladd, 2008, p. 107</xref>, where his (3) can be interpreted in this way), although it figures prominently in the description of Dutch, which has a similar intonation system to English (<xref ref-type="bibr" rid="B17">Collier &amp; &#8217;t Hart 1980</xref>; <xref ref-type="bibr" rid="B62">&#8217;t Hart et al., 1990, p. 116</xref>). Panel (a) in Figure <xref ref-type="fig" rid="F9">9</xref> gives the <italic>F<sub>0</sub></italic> and speech waveform of an English example. It contrasts with the contour in panel (b), which would appear to have an IP-boundary after <italic>finance committee</italic> (<xref ref-type="bibr" rid="B31">Gussenhoven, 2004, p. 305</xref>; see also panels (c) and (d)). Section 4 reports an experiment that was designed to decide whether a medial boundary exists in contour (a).</p>
<fig id="F9">
<label>Figure 9</label>
<caption>
<p>A sharp pre-nuclear fall on <italic>fi-</italic> in <italic>But the FInance committee needn&#8217;t be inVOLVED in this</italic> (panel a) and a contour with an IP-boundary after <italic>finance committee</italic> (panel b). Stylized versions are given in panels (c) and (d). Male GBE speaker. This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav9a">http://dx.doi.org/10.5334/labphon.30.wav9a</ext-link> and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav9b">http://dx.doi.org/10.5334/labphon.30.wav9b</ext-link></p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74556/"/>
</fig>
</sec>
<sec>
<title>5.2 A perception experiment</title>
<p>Adverbs like <italic>honestly</italic> and <italic>oddly</italic> can modify adjectives, predicates, and clauses. Only in the third case are they obligatorily separated from the clause by an intonational boundary. This is illustrated in (16a), which minimally contrasts with (16b), where <italic>honestly</italic> modifies a predicate.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(16)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>{She treated him}{honestly}&#8216;I am honest when telling you that she treated him&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>{She treated him honestly} &#8216;The way she treated him was honest&#8217;</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>In order to find evidence for the assumption that the interpretation of sentence-final English adverbs depends on the presence of an intonational boundary before the adverb, more specifically that contour (a) of Figure <xref ref-type="fig" rid="F9">9</xref> does not have an internal intonational boundary, a semantic judgement task was used in which native speakers of English identified one of two meanings of string-identical sentences of the kind illustrated in (16) which had been provided with a number of artificial <italic>F<sub>0</sub></italic> contours.</p>
</sec>
<sec sec-type="methods">
<title>5.2.1 Method</title>
<p>Four minimal sentence pairs with string-ambiguous adverbials were composed (17).</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(17)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>She TREATED the poor man(,) HONESTLY</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>I THOUGHT she responded(,) ODDLY</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>c.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>He NEVER acted(,) STRANGELY</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>d.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>He DEALT with the woman(,) HONESTLY</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>The eight sentences were recorded by a female native speaker of MAE in her thirties from Portland, Oregon. By judiciously cutting and pasting sections in the stretch of the waveform before the adverbial, one durational hybrid of each pair of utterances was created, using the software Praat (<xref ref-type="bibr" rid="B8">Boersma &amp; Weenink, 1992&#8211;2009</xref>). The single-IP versions were used as the source utterance in the case of (17a, c) and the split-IP ones in the case of (17b, d). Appendix C gives the durations of the sections in the original speech files for two sentences with and without boundary whose averaged durations were created in the hybridized source files. By using these as source utterances for <italic>F<sub>0</sub></italic> manipulation, we neutralized the effect of any durational marking of the IP-boundary in the original recordings. With the help of the resynthesis program in Praat (<xref ref-type="bibr" rid="B8">Boersma &amp; Weenink, 1992&#8211;2009</xref>), we then superimposed 12 declining <italic>F<sub>0</sub></italic> contours on each of these four sound files, with <italic>F<sub>0</sub></italic> values which are representative of the speaker&#8217;s original utterances (see Table <xref ref-type="table" rid="T3">3</xref> for these values; unmarked turning points have the same values as equivalent points with <italic>F<sub>0</sub></italic>-labels). The twelve contours come in two sets of six, as shown in the six cells of Table <xref ref-type="table" rid="T3">3</xref>.</p>
<table-wrap id="T3">
<label>Table 3</label>
<caption>
<p>Schematic representations of double and single IPs for three two-accent contours with &#8216;Fall-rise&#8217;, &#8216;Fall&#8217;, and &#8216;High rise&#8217; pitch accents for the first accent and falling pitch accents on the second accent, with F0 values (Hz) for turning points as used in the artificial contours. The interrupted line indicates versions of the contour with initial %H.</p>
</caption>
<table>
<tr>
<th align="left"></th>
<th align="left">With medial IP-boundary</th>
<th align="left">Without medial IP-boundary</th>
</tr>
<tr>
<td colspan="3">
<hr/></td>
</tr>
<tr>
<td align="left">Fall-rise</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74557/"/></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74558/"/></td>
</tr>
<tr>
<td align="left">Fall</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74559/"/></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74560/"/></td>
</tr>
<tr>
<td align="left">High rise</td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74561/"/></td>
<td align="left"><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74562/"/></td>
</tr>
</table>
</table-wrap>
<p>In order to increase the variation in the stimuli, one set had a low-pitched syllable before the first pitch accent (<italic>She</italic>, <italic>I</italic>, <italic>He</italic>, <italic>He</italic>), while these syllables had high pitch in the other set, as indicated by the interrupted sections, phonologically equivalent to initial %H. The crucial comparison is that between the contours in the two cells of the row labelled &#8216;Fall-rise&#8217;, which reproduce the contrast in Figure <xref ref-type="fig" rid="F9">9</xref>. As a baseline, we included a contour with an accent-lending fall which <italic>does</italic> signal an intonational boundary in other descriptions of English, as shown in the row labeled &#8216;Fall&#8217;. The pitch after the first <italic>F<sub>0</sub></italic> peak continues low; its counterpart without an intonational boundary is taken to have a slow fall between the accent peaks. As a further control, the contours given under &#8216;High rise&#8217; were included. Here, the fall just before the rise towards the second peak is also taken to predict an intonational boundary. The counterpart without the boundary has high level pitch between the accent peaks. It is stressed that the <italic>F<sub>0</sub></italic> manipulations were applied to only four soundfiles, one for each sentence, and that any effects are therefore based on <italic>F<sub>0</sub></italic> differences only. The interrupted contour sections correspond to the implied IP-boundary.</p>
<sec>
<title>5.2.2. Procedure</title>
<p>Contours were exhaustively paired within each set of six contours for each of the four source files, excluding pairings of identical contours. This gave two sets of 30 pairs, one with low and one with high beginnings. In order to avoid an unmanageably large set of stimuli, which would arise if we had included 30 (pairings) &#215; 2 (sets) &#215; 4 (sentences) = 240 stimulus pairs, we composed two sets of 30 stimuli, one with initial low <italic>F<sub>0</sub></italic> selected from sentences (17b, d) and one with initial high <italic>F<sub>0</sub></italic> selected from sentences (17a, c) (see Appendix A). The inclusion of all four source files was intended to avoid fatigue and boredom among the participants. Two test versions were prepared with counterbalanced orders of these 60 stimulus pairs, augmented with four filler pairs inserted at the beginning. Moreover, the members of the stimulus pairs occurred in reversed order in the two test versions.<xref ref-type="fn" rid="n7">7</xref></p>
<p>Seventeen native speakers of American English, approximately equally divided over male and female genders, participated in this semantic identification task. Fifteen participants were recruited from the student population of the Linguistics Department of UC Berkeley, while two were staff members in similar departments in the UK and the Netherlands. Each stimulus pair was presented once, with a latency of 800 ms after a warning signal. The interval between the members of each pair was 800 ms, while 5 seconds elapsed between each pair and the warning signal for the next pair. The participants, 8 of whom did one test version and 9 the other, were asked to identify which of the two members in each pair corresponded best with the interpretation of the sentence-final adverb as a predicate modifier (Version A) or a sentence modifier (Version B; see Appendix B for these instructions). They gave their judgements on a 3-point scale, labelled &#8216;1&#8217; (for the first member), &#8216;0&#8217; (for no preference) and &#8216;2&#8217; (for the second member).</p>
</sec>
<sec>
<title>5.2.3. Results</title>
<p>The 1, 0, 2 score values were converted to &#8211;1, 0, +1 (version A) and +1, 0, &#8211;1 (version B), respectively, so that a higher score represents a higher degree of predicate adverb interpretation of the adverb. A RM Anova on the scores pooled over source files was performed with Initial Boundary Tone, Medial Boundary, and First Pitch Accent as factors. It only showed significant main effects for Medial Boundary (<italic>F</italic><sub>2,16</sub> = 424,254; <italic>p</italic> &lt; 0.0001) and First Pitch Accent (<italic>F</italic><sub>1.621,16</sub> = 134,797; <italic>p</italic> &lt; 0.0001; Huynh-Feldt corrected). Since there was no effect of the <italic>F<sub>0</sub></italic> of the contour beginning, scores were averaged over low-pitched and high-pitched initial syllables and displayed in Figure <xref ref-type="fig" rid="F10">10</xref>. Post-hoc pairwise comparisons show that the High-rise pitch accent attracted significantly higher scores for the interpretation as a predicate adverb than both the Fall (<italic>p</italic> &lt; 0.01) and the Fall-rise (<italic>p</italic> &lt; 0.001).</p>
<fig id="F10">
<label>Figure 10</label>
<caption>
<p>Perceived scores for predicate adverb, aggregated over two test versions and high and low beginning stimuli for three contours with (right) and without (left) a medial IP- boundary.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74563/"/>
</fig>
</sec>
<sec>
<title>5.2.4. Discussion</title>
<p>The results confirm the interpretation of the contours in the second column in Table <xref ref-type="table" rid="T3">3</xref> as having no medial intonational boundary. Crucially, the Fall-rise contour in the column &#8216;With medial IP-boundary&#8217; is interpreted to differ from the Fall-rise contour in the second in the same way as do the single-IP and two-IP versions of the Fall and High-rise contours. There is therefore no motivation for a transcription of the first Fall-rise contour with L- after the first pitch accent.</p>
<p>Three additional points are made. First, the finding that the High-rise contours are more readily interpreted as lacking an intonational boundary than either the Fall-rise or Fall contours is attributed to the low phonetic salience of the <italic>F<sub>0</sub></italic> features separating the two pitch accents. In the contour without medial boundary, the pitch continues level from one peak to the next, <italic>modulo</italic> the declination, and for the contour with the medial IP-boundary, it is only the falling-rising pitch movement just before the adverb which can be held responsible for the perceptual effect of the intonational boundary. Second, it is striking that this subtle phonetic feature has the same interpretation effect as the more substantial phonetic differences between the two contours for the pre-boundary Fall-rise and the Fall. The fact that there is no interaction between Medial Boundary and First Pitch Accent means that the effect sizes of the medial boundary do not vary across the three contour types. There is therefore no evidence in these data for two intonational prosodic constituents, like the intermediate phrase in the case of the right-hand Rise and Fall contours, and the intonational phrase in the case of the right-hand Fall-rise contour. Thirdly, the absence of any effect of %H was to be expected, as it has no role to play in signalling an upcoming boundary.</p>
<p>These results replicate those obtained in Gussenhoven (<xref ref-type="bibr" rid="B32">2008</xref>) for Dutch. In that experiment, participants indicated their interpretation of three ambiguous words on a 5-point scale, which had a modal adverb at one end and a predicative adjective at the other. There were three such words, one example being <italic>vast</italic>. As a modal adverb it means &#8216;surely&#8217;, as in <italic>Ze zit VAST op de SNELweg</italic> &#8216;She must surely be on the motorway&#8217;, while the predicative adjective means &#8216;stuck&#8217;, giving &#8216;She has got stuck on the motorway&#8217;. If the pitch accent on the target word, here <italic>VAST</italic>, is identical to that on the VP (here <italic>zit op de SNELweg</italic>) and there is an IP-boundary between them, a pattern arises that is referred to as &#8216;tone concord&#8217; by Wells (<xref ref-type="bibr" rid="B69">2006, p. 85</xref>) and which uniquely gives the interpretation of predicative adjective. However, in the interpretation as a modal adverb, there is no IP-boundary. Ignoring details, those results were the same as those reported here for English.</p>
</sec>
<sec>
<title>5.2.5. The interpretation of the prenuclear fall-rise</title>
<p>According to the exposition so far, neither MAE_ToBI nor the off-ramp analysis can account for the results for the Fall-rise contours. In the off-ramp analysis, a pre-nuclear fall is described as H*L, but this would rather give a slow fall, not a sharp fall plus a slow rise. It is reasonable to assume that a historical reinterpretation of {%L H*L H%}{%L H*L L%} as a single IP retained the salient medial H% at the expense of medial %L. If this H-tone is reinterpreted as the final tone in a tritonal prenuclear pitch accent, as in {%L H*LH H*L L%}, the realization with H in rightmost position follows from the grammar. It will locate the target of the final trailing tone just before the target of the next H*, and interpolate to it from the target of preceding L (cf. <xref ref-type="bibr" rid="B18">Cruttenden, 1997, p. 76</xref>). This contour is presented by O&#8217;Connor &amp; Arnold (<xref ref-type="bibr" rid="B48">1973</xref>), here given as (18), though analyzed there as a contour containing an IP boundary. Figure <xref ref-type="fig" rid="F11">11</xref> gives the pitch track of their recorded example, overlaid with a resynthesized version, which to my ear sounds the same. The actual phrasing of this contour is somewhat ambiguous due to the long duration of the final syllable of <italic>Paris</italic>, which suggests a pronunciation with two IPs.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(18)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>(The food in) \/Paris was su \perb</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<fig id="F11">
<label>Figure 11</label>
<caption>
<p><italic>F<sub>0</sub></italic> track (black speckles) with superimposed smoothed contour (grey speckles) of <italic>Paris was superb</italic> (speech file from <xref ref-type="bibr" rid="B48">O&#8217;Connor &amp; Arnold, 1973</xref>). This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav11a">http://dx.doi.org/10.5334/labphon.30.wav11a</ext-link> and <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wav11b">http://dx.doi.org/10.5334/labphon.30.wav11b</ext-link>.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74564/"/>
</fig>
<p>The analysis of (16a, b) in the off-ramp view is shown in (19a, b). As observed above, the IP-final H% of (19a) ends up as a third tone in the prenuclear pitch accent in (19b), which aligns rightmost, as usual. The initial %L in the second IP is deleted in the restructured form. An unexpected confirmation of the analysis in (19b) for Dutch, where the same contours exist, is provided by &#8217;t Hart et al. (<xref ref-type="bibr" rid="B62">1990</xref>), who reported an accelerated rise following the slow rise, occurring just before the second accented syllable, which they labeled &#8216;5&#8217; (see panel (c) in Figure <xref ref-type="fig" rid="F9">9</xref>). Similarly, Steedman (<xref ref-type="bibr" rid="B60">2014</xref>) discusses this contour in terms of how the theme is signaled, placing the intonational boundary between the theme <italic>Anna will marry</italic> (pronounced L+H* LH%) and the rheme <italic>Manny</italic> (pronounced H* LL%, his example (10)).</p>
<p>It is tempting to interpret the two consecutive high targets in these descriptions as reflecting the targets of prenuclear trailing H and nuclear H*, respectively.</p>
<p>MAE_ToBI cannot easily account for this contour. A newly introduced prenuclear H*+L would have the arbitrary property of requiring a nuclear pitch accent beginning with a H-tone, to make sure there is a slow rise from the prenuclear accented syllable. This measure would however not account for the wider facts, since pre-nuclear H*LH may also appear before pitch accents beginning with L*, in which case there would be no H-tone to explain the slow rise (<xref ref-type="bibr" rid="B29">Gussenhoven, 1983, p. 63</xref>). The alternative decision to introduce a pre-nuclear H*+L+H would have the disadvantage of requiring a unique timing policy for the final H tone, in order to prevent it from being realized immediately after the pitch fall described by H*+L. In other words, while the off-ramp analysis can naturally incorporate a pre-nuclear H*LH, the on-ramp analysis cannot.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(19)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74565/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx19a">http://dx.doi.org/10.5334/labphon.30.wavEx19a</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74566/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx19b">http://dx.doi.org/10.5334/labphon.30.wavEx19b</ext-link></p>
<p>The contours labeled &#8216;Fall&#8217; have the representations in (20a, b), those labeled &#8216;Rise&#8217; are given in (21a, b). As will be clear, the a-examples all have the same phonological boundary, a prediction that was supported by the results of the perception experiment.</p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(20)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74567/"/></p></list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx20a">http://dx.doi.org/10.5334/labphon.30.wavEx20a</ext-link></p>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74568/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx20b">http://dx.doi.org/10.5334/labphon.30.wavEx20b</ext-link></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(21)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74569/"/></p></list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx21a">http://dx.doi.org/10.5334/labphon.30.wavEx21a</ext-link></p>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74570/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>This audio content is available at: <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.wavEx21b">http://dx.doi.org/10.5334/labphon.30.wavEx21b</ext-link></p>
</sec>
</sec>
</sec>
<sec>
<title>6 Other empirical evidence</title>
<p>The identification of the falling section of an <italic>F<sub>0</sub></italic>-peak as a pitch accent would appear to avoid the cases of underanalysis and overanalysis by MAE_ToBI which were discussed in Sections 3, 4, and 5. Two findings have been presented that more specifically support the off-ramp view. First, Dilley et al. (<xref ref-type="bibr" rid="B21">2005</xref>) show that there is a low correlation between the timings of the first valley and the peak in <italic>F<sub>0</sub></italic> rise-falls, suggesting that the targets of L and H* do not obey a constant interval, as suggested by the MAE_ToBI L+H* pitch accent, but are timed independently with reference to the segmental string. Conversely, Barnes et al. (<xref ref-type="bibr" rid="B2">2010</xref>) show that the target of the L-tone <italic>after</italic> H* is located with reference to the target of H*, and not with reference to any following segmental landmark, which does not support the MAE_ToBI analysis of the fall as being composed of H* followed by a heteromorphemic phrase tone. The latter result was also obtained for a number of varieties of continental West Germanic (<xref ref-type="bibr" rid="B50">Peters et al., 2015</xref>). These two sets of findings are just as would be expected under an off-ramp view, in which the rise is defined by heteromorphemic tones and the fall by tautomorphemic tones. In addition to these alignment facts, there are pitch span effects for Dutch that appear to confirm the off-ramp view. Chen (<xref ref-type="bibr" rid="B13">2011</xref>) measured the pitch span of rises and falls of accentual pitch peaks on the S of SVO sentences in elicited adult speech. In about half the data, the S was contextually focused, while in the remainder it was topic, the O being focused. When dividing the data up into utterances in which the pitch after H* continued at a high level and utterances in which the pitch sloped down from the peak, she found that the H*+level contours differed significantly in the span of the rise towards H* as a function of the focus structure, rise spans being wider because of a lower end point. However, the rises in the H*+fall were not significantly different in the two focus conditions; rather, it was the fall that had a significantly wider pitch span, because it ended lower in the focus condition. These results do not match the on-ramp analysis, which would describe both H*-peaks as consisting of a pitch accent that represents the rise, L+H*. By contrast, the off-ramp analysis analyzes the H*+level as H* (preceded by a %L boundary tone), while the H*+fall is analyzed as H*L. Focus in Dutch can thus coherently be described as causing a raising of H* and a hyperarticulation of the fall represented by H*L.</p>
<p>Lastly, it is reiterated here that the results of Gussenhoven &amp; Rietveld (<xref ref-type="bibr" rid="B33">1991</xref>) favoured the off-ramp analysis of Gussenhoven (<xref ref-type="bibr" rid="B29">1983</xref>) over the analysis in Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>). The two sets of 210 differences in terms of phonological elements among 15 nuclear melodies as expressed in those two theories showed a modest correlation of <italic>r</italic> = 0.38, meaning that the theories made very different predictions about the degree of similarity between pairs of nuclear melodies. Semantic differences obtained from a perception experiment with auditory stimuli representing those same pairs of nuclear melodies correlated fairly well with the off-ramp theory (<italic>r</italic> = 0.57), while no significant correlation was found between the Pierrehumbert data and the perception data.</p>
</sec>
<sec>
<title>7 Summary and conclusion</title>
<p>The off-ramp intonation grammar derived above and earlier provided in Gussenhoven (<xref ref-type="bibr" rid="B31">2004, p. 313</xref>)<xref ref-type="fn" rid="n8">8</xref> is summarized in (22), with the conventions in (23).<xref ref-type="fn" rid="n9">9</xref></p>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(22)</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p><graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="/article/id/6178/file/74571/"/></p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>(23)</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>I.</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>The last trailing tone of a prenuclear pitch accent aligns rightmost.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Other trailing tones align leftmost.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>II.</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Within a pitch accent, interpolations are linear.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Otherwise, unspecified speech is governed by the leftmost tone.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>III.</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>a.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Within a pitch accent, downstep of H after H is obligatory.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<list list-type="gloss">
<list-item>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>&#160;</p></list-item>
</list>
<list list-type="wordfirst">
<list-item><p>b.</p></list-item>
</list>
</list-item>
<list-item>
<list list-type="sentence-gloss">
<list-item>
<list list-type="final-sentence">
<list-item><p>Otherwise, downstep of H* is optionally triggered by a preceding H.</p></list-item>
</list>
</list-item>
</list>
</list-item>
</list>
<p>Significantly, the conventions in (23) refer to pitch accents, as opposed to similar tone sequences belonging to different morphemes. The off-ramp analysis thus brings out the phonological and morphological relevance of this concept, making its tones distinct from otherwise identical sequences of tones. This sensitivity to the morphemic structure strengthens the case of the off-ramp analysis, because reference to morphological structure is a routine feature of phonological generalizations across languages.</p>
<p>It was suggested that a historical accident, Pierrehumbert&#8217;s (<xref ref-type="bibr" rid="B51">1980</xref>) adoption of an equivalent of the focus-marking H from Bruce&#8217;s (<xref ref-type="bibr" rid="B12">1977</xref>) tonal phonology of Central Swedish, lay behind her decision to analyze an accent-marking rising-falling pitch configuration in Mainstream American English as a rising pitch accent L+H* followed by a low tone from some other source, instead of a falling pitch accent H*+L preceded by a low tone from some other source. This on-ramp analysis led to a number of questionable properties of her analysis, many of which were inherited by a widely used transcription system for the language, MAE_ToBI. First, it created the need for two further tones after a nuclear pitch accent, later leading to the introduction of a tonally marked prosodic constituent, the intermediate phrase, by the side of the higher-ranking intonational phrase (<xref ref-type="bibr" rid="B6">Beckman &amp; Pierrehumbert, 1986</xref>). Since no other analysis of a West Germanic language had earlier seen the need for that constituent,<xref ref-type="fn" rid="n10">10</xref> the MAE_ToBI intonational phrasing analysis is unique among those many analyses of West Germanic intonation. Other assumptions which were in part generated by the on-ramp view and which were questioned here include the sagging of pitch between H*-targets, instead of sustained high pitch; the use of leading H in H+H* to ensure continued high pitch preceding downstepped !H*-targets, which usurps a general function of leading H to describe pre-accentual peaks; downstepped H-tones other than !H*, instead of downstep of H* only; linear interpolations between pitch accents, instead of a continuation of the left-hand tonal target; and the equation of L*-prefixed (&#8216;scooped&#8217; or &#8216;delayed&#8217;) contours with rising pitch accents.</p>
<p>Section 2 gave a summary statement of the MAE_ToBI analysis. Section 3 inventorized cases of underanalysis, the absence of a transcription for some contour, and Section 4 did the same for cases of overanalysis, the existence of more than one transcription for some contour. Section 5 presented perception data that suggest that MAE_ToBI&#8217;s prediction of an intonational boundary is false in the case of a prenuclear sharp fall which is followed by a gradual rise to the next accented syllable. Those data also revealed a lack of evidence for a two-tier intonational phrasing structure.</p>
<p>Throughout the discussion, it was shown how the opposite choice, the identification of a falling pitch accent in the accent-lending rise-fall (an off-ramp analysis), avoids the disadvantages of the MAE_ToBI analysis. The off-ramp analysis was similarly a historical accident, since it tacitly continued the off-ramp view of the British tradition (<xref ref-type="bibr" rid="B29">Gussenhoven, 1983</xref>, <xref ref-type="bibr" rid="B31">2004</xref>). It shares with MAE_ToBI the incorrect prediction of a phrase break as described in Section 5. In the off-ramp case, this is because any trailing L-tone in a pre-nuclear pitch accent will be realized late, creating a slow fall rather than a slow rise. However, it was argued that the introduction of a tritonal pre-nuclear pitch accent H*LH, which was claimed to have resulted from a phonological change triggered by phrasal restructuring, fits neatly into the tone grammar that was independently yielded by the off-ramp view. In addition, two potential cases of overanalysis were identified for the off-ramp analysis. One concerned the occurrence of a L*H pitch accent by the side of a L*-prefixed set of nuclear pitch accents beginning with H*. The second was the generation of a set of contours with final truncated falls, &#8216;slumps&#8217;, which have not been attested in MAE and would probably be considered alien to that dialect if presented to its speakers. It is to be noted, however, that in both cases the overgeneration concerns identifiably different contours from other contours generated by the grammar, which was not true for the overgeneration of representations in MAE_ToBI. In the first case, more empirical evidence is required to validate the distinction predicted by the off-ramp analysis between L*H and H* prefixed by L*. To cover the second class of contours, we have appealed to a wider coverage of the grammar than that for any specific variety of English, such that varieties may fail to use contours that are legitimate products of the grammar. Varieties are known in any event to differ in the frequency of use of contours (Section 3.5), and a stochastic structure as envisaged by Dainora (<xref ref-type="bibr" rid="B20">2006</xref>) may be a goal of future research.</p>
<p>The above suggestion of a grammar which serves a group of closely related varieties of a language is not intended to blur the fact that we exclusively evaluated a phonological analysis of English, MAE_ToBI, and compared it with an alternative analysis. That is, there is no direct implication that analyses of other languages should be revised in similar ways. Phonological diversity is likely to apply to intonational structure as much as it does to segmental structure. A two-level intonational phrasing structure of the type that was introduced by Beckman and Pierrehumbert (<xref ref-type="bibr" rid="B6">1986</xref>) appears to be well-motivated in the case of varieties of Bengali (<xref ref-type="bibr" rid="B35">Hayes &amp; Lahiri, 1991</xref>; <xref ref-type="bibr" rid="B36">Kahn, 2014</xref>), to give just one example. On-ramp and off-ramp analyses appear to apply to similar rising-falling contours in different Romance languages (<xref ref-type="bibr" rid="B22">Frota, this issue</xref>). Empty space between a nuclear pitch accent and an IP-final boundary tone is pronounced with left-aligned targets of the boundary tone in the tonal dialect of Roermond Dutch, but with right-aligned tones of the pitch accent in non-tonal Dutch (<xref ref-type="bibr" rid="B30">Gussenhoven, 2000</xref>, <xref ref-type="bibr" rid="B31">2004</xref>), and so on. More empirical research into issues of the phonological representation of intonation is a desideratum. Pierrehumbert&#8217;s (<xref ref-type="bibr" rid="B51">1980</xref>) conceptualization of the difference between phonological structure and phonetic implementation will provide an important background here, given that many communicative effects of pitch variation are non-structural, i.e., paralinguistic (<xref ref-type="bibr" rid="B40">Ladd, 2008, p. 34</xref>).</p>
</sec>
</body>
<back>
<fn-group>
<fn id="n1">
<p>Riad (<xref ref-type="bibr" rid="B55">2014, p. 254</xref>) analyses the Central Swedish lexical tone distinction privatively, such that the intonational tones only appear after the lexical tone in the case of Accent 2.</p>
</fn>
<fn id="n2">
<p>&#8217;t Hart et al. (<xref ref-type="bibr" rid="B62">1990</xref>) replaced the accent-lending falling movement (&#8216;A&#8217;) with a movement that had been used to describe the transition between a high-ending and a low-beginning IP (&#8216;B&#8217;), thus opting for an on-ramp analysis in the summary of their work.</p>
</fn>
<fn id="n3">
<p>A reviewer pointed out that definitions of &#8216;core&#8217; or &#8216;abstract&#8217; (<xref ref-type="bibr" rid="B18">Cruttenden, 1997</xref>) meanings for intonational morphemes may be problematic and that factors like politeness, social distance, and physical distance need to be considered for a better understanding of intonational meaning. For instance, B&#243;rras-Comes et al. (<xref ref-type="bibr" rid="B10">2015</xref>) show that compared to non-chanted vocatives, the vocative chant of Central Catalan is favoured by larger physical distance and speaker superiority. These factors may well also apply to the English vocative chant..</p>
</fn>
<fn id="n4">
<p>Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>) had H+L* instead of MAE_ToBI&#8217;s H+ !H*, with a convention that L* in H+L* was realized as !H*.</p>
</fn>
<fn id="n5">
<p>Until Gussenhoven et al. (<xref ref-type="bibr" rid="B34">1999&#8211;2003</xref>), tonally unspecified stretches of speech between pitch accents were filled by interpolations instead of double alignment. Grice et al. (<xref ref-type="bibr" rid="B28">2009</xref>) assumed linear interpolation in their critique of my off-ramp analyses of (2004, <xref ref-type="bibr" rid="B27">2005</xref>), which however include &#8216;continuation&#8217;. The term &#8216;spreading&#8217; was avoided because of its implication of tonal association. The relevance of interpolation shapes for the perception of the alignment of pitch targets was demonstrated by Barnes, Veilleux, Brugos, and Shattuck-Hufnagel (<xref ref-type="bibr" rid="B2">2010</xref>, <xref ref-type="bibr" rid="B3">2012</xref>), who introduced a Tonal Center of Gravity (TCG) to predict perceived alignment. This paper recognizes this view of pitch movements, but it is not directly relevant for the discussion of the more coarse-grained alignments of tones in this article.</p>
</fn>
<fn id="n6">
<p>MAE_ToBI has two ways of indicating prosodic boundaries. In addition to the tonal one, there are the &#8216;break indices&#8217;, whereby digits 0 and 1 indicate a clitic group boundary and a word boundary, respectively, and 3 and 4 an ip and an IP boundary, respectively. Digit 2 is there for cases where the tonal transcription implies boundaries that are not perceived to be as strong or as weak as the transcriber&#8217;s tonal transcription is felt to imply. The distinction between 0 and 1 falls outside our interest. The information about perceived outliers in pause durations, the only non-redundant information in the break indices, is not relevant to our research question.</p>
</fn>
<fn id="n7">
<p>A typing error caused one of the 30 pairs to be missing, for which aggregate scores were imputed, and another to appear twice, for which scores were averaged. For details, see Appendix A.</p>
</fn>
<fn id="n8">
<p>My first attempt at an autosegmental analysis was based on Goldsmith (<xref ref-type="bibr" rid="B24">1980</xref>) and a familiarity with &#8217;t Hart &amp; Collier (<xref ref-type="bibr" rid="B61">1980</xref>) and O&#8217;Connor &amp; Arnold (<xref ref-type="bibr" rid="B48">1973</xref>), none of which explicitly featured boundary tones (cf. <xref ref-type="bibr" rid="B39">Ladd, 1983</xref>). My 1983 description of English was based on three pitch accents that could undergo modifications, much as in Ladd (<xref ref-type="bibr" rid="B37">1978</xref>, <xref ref-type="bibr" rid="B39">1983</xref>), and treated the effects of boundary tones as the phonetic realizations of the pitch accents. In that description, I assumed that trailing tones of nuclear pitch accents &#8216;spread&#8217; (i.e., &#8216;continue&#8217; in the terminology used in this paper), but recoiled from assuming that final tones of prenuclear pitch accents do so (<xref ref-type="bibr" rid="B39">1983, p. 72</xref>), instead opting for a Tone Linking Rule which deleted trailing tones of pre-nuclear pitch accents. The generalized notion of continuation was originally formulated for Dutch (<xref ref-type="bibr" rid="B34">Gussenhoven et al., 1999&#8211;2003</xref>). An English version appeared as chapter 15 in Gussenhoven (<xref ref-type="bibr" rid="B31">2004</xref>). Unlike the wider formulation there, which maintained the 1983 proposal of a modification [DELAY] for both L*-initial and H*-initial pitch accents, (23) allows the affixation of the L*-prefix to H* only. This agrees with Cruttenden (1986, p. 123).</p>
</fn>
<fn id="n9">
<p>The number of contours (22) generates is larger than that for MAE_ToBI, and <italic>ceteris paribus</italic> cases of underanalysis should be rarer, while the risk of overanalysis might be expected to be higher. Prefix L*, which may be attached to H*, H*L, and H*H, puts the number of nuclear contours at 2 (H-Prefix) &#215; 8 (5 + 3 L*-prefixed pitch accents) &#215; 3 (IP-endings) or 48 nuclear melodies. With 2 IP-beginnings and 5 prenuclear pitch accents, this gives 480 two-accent contours. Here, my assumption is that prenuclear scooped contours typically imply nuclear scooped contours, so that L*-prefixation is not counted separately for prenuclear accents. Because downstep is obligatory on H* after leading H on H*, I assume there are no additional downstepped versions of contours with leading H. Among %L-beginning contours, four pre-nuclear pitch accents have a H-tone (H*, H*L, H*LH, L*H), while six nuclear pitch accents have a targetable H* (i.e., not preceded by a leading H) (H*, H*L, H*+H, L*=H*, L*=H*L, and L*=H*+H), i.e., 4 &#215; 6 &#215; 3 (IP-endings) or 72 downstepped contours, 144 if H% beginning ones are included. Among the %H-beginning contours which have one H* in either prenuclear or nuclear position, there are 18 with L* in prenuclear position before a nuclear pitch accent with H* but without leading H, and 18 with L* or L*H in nuclear position with a prenuclear pitch accent containing H*, adding another 36 downstepped contours, or 660 in all.</p>
</fn>
<fn id="n10">
<p>This is not to say that no phrasal distinctions have been made in earlier descriptions. &#8216;Subordinated&#8217; tone groups have been claimed to account for IPs with reduced pitch range (<xref ref-type="bibr" rid="B19">Crystal, 1969, p. 244</xref>), while Trim (<xref ref-type="bibr" rid="B64">1959</xref>) pleads for a continued use of double bar (||) and single bar (|) boundary markers to imply freedom of tonal dependence across the double bar. In modern terms, these correspond to utterance and IP boundaries, respectively. Pierrehumbert (<xref ref-type="bibr" rid="B51">1980</xref>), among others, incorporated utterance-final unaccented IPs in her analysis, following Bing&#8217;s (<xref ref-type="bibr" rid="B7">1979</xref>) &#8216;O-domains&#8217;. None of these phrasal distinctions concern MAE_ToBI&#8217;s ip, however.</p>
</fn>
</fn-group>
<sec>
<title>Supplementary Files</title>
<p>For accompanying TextGrid, Pitch, and wav files, go to <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="http://dx.doi.org/10.5334/labphon.30.smo">http://dx.doi.org/10.5334/labphon.30.smo</ext-link></p>
</sec>
<ack>
<title>Acknowledgements</title>
<p>I am indebted to audiences for their comments on the semantic identification experiments for English and Dutch at the workshop on Transcription of Intonation in the Ibero-Romance Languages at the 7<sup>th</sup> PaPI conference in Braga (<italic>Predicting boundaries from ToDI transcriptions</italic>, 26 July 2007), the Linguistics Colloquium at Radboud University Nijmegen, the 8th Phonetics Conference of China (PCC 2008) in Beijing (<italic>Evidence for ToDI from semantic judgements</italic>, 18&#8211;20 April 2008), and the poster session at Speech Prosody 2012 in Rio de Janeiro (<italic>Semantic judgments as evidence for the intonational structure of Dutch</italic>, August 2008). I am grateful to Sam Tilsen for his help in running the perception experiment reported in Section 3 in Berkeley, to Mybeth Lahey for her help with processing the scores, to Joop Kerkhoff for technical assistance, to James McQueen, Scott Moisik, Brigitte Planken, Natasha Warner, and Anne Wichmann for recording examples, to Jos&#233; Hualde for discussion, and to Martine Grice, Bob Ladd, J&#246;rg Peters, and two anonymous reviewers for their comments on earlier versions.</p>
</ack>
<sec>
<title>Competing Interests</title>
<p>The author declares that they have no competing interests.</p>
</sec>
<ref-list>
<ref id="B1">
<label>1</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Analytical Decisions in Intonation Research and the Role of Representations: Lessons from Romani</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>6</elocation-id>
<fpage>1</fpage>
<lpage>43</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.14</pub-id>
</element-citation>
</ref>
<ref id="B2">
<label>2</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Barnes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Veilleux</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Brugos</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Turning points, tonal targets, and the English L- phrase accent</article-title>
<source>Language and Cognitive Processes</source>
<year iso-8601-date="2010">2010</year>
<volume>25</volume>
<fpage>982</fpage>
<lpage>1023</lpage>
<pub-id pub-id-type="doi">10.1080/01690961003599954</pub-id>
</element-citation>
</ref>
<ref id="B3">
<label>3</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Barnes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Veilleux</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Brugos</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Tonal center of gravity: A global approach to tonal implementation in a level-based intonational phonology</article-title>
<source>Laboratory Phonology</source>
<year iso-8601-date="2012">2012</year>
<volume>3</volume>
<fpage>337</fpage>
<lpage>383</lpage>
<pub-id pub-id-type="doi">10.1515/lp-2012-0017</pub-id>
</element-citation>
</ref>
<ref id="B4">
<label>4</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Ayers</surname>
<given-names>G. M.</given-names>
</name>
</person-group>
<source>Guidelines for ToBI labeling</source>
<year iso-8601-date="1994">1994</year>
<comment>Retrieved from <uri>http://www.speech.cs.cmu.edu/tobi/ToBI.0.html</uri></comment>
</element-citation>
</ref>
<ref id="B5">
<label>5</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>The original ToBI system and the evolution of the ToBI framework</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005">2005</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>9</fpage>
<lpage>45</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0002</pub-id>
</element-citation>
</ref>
<ref id="B6">
<label>6</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
</person-group>
<article-title>Intonational structure of English and Japanese</article-title>
<source>Phonology Yearbook</source>
<year iso-8601-date="1986">1986</year>
<volume>3</volume>
<fpage>255</fpage>
<lpage>309</lpage>
<pub-id pub-id-type="doi">10.1017/S095267570000066X</pub-id>
</element-citation>
</ref>
<ref id="B7">
<label>7</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>Bing</surname>
<given-names>J. M.</given-names>
</name>
</person-group>
<source>Aspects of English prosody, (Unpublished doctoral dissertation)</source>
<year iso-8601-date="1979">1979</year>
<publisher-name>University of Massachusetts</publisher-name>
</element-citation>
</ref>
<ref id="B8">
<label>8</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Boersma</surname>
<given-names>P.</given-names>
</name>
<name>
<surname>Weenink</surname>
<given-names>D.</given-names>
</name>
</person-group>
<source>Praat: doing phonetics by computer</source>
<year iso-8601-date="1992&#8211;2009">1992&#8211;2009</year>
<comment>Retrieved from <uri>http://www.praat.org</uri></comment>
</element-citation>
</ref>
<ref id="B9">
<label>9</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bolinger</surname>
<given-names>D.</given-names>
</name>
</person-group>
<article-title>Intonation: Levels vs. configurations</article-title>
<source>Word</source>
<year iso-8601-date="1951">1951</year>
<volume>14</volume>
<fpage>109</fpage>
<lpage>149</lpage>
<pub-id pub-id-type="doi">10.1080/00437956.1958.11659660</pub-id>
</element-citation>
</ref>
<ref id="B10">
<label>10</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Borr&#225;s-Comes</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Sichel-Bazin</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Prieto</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>Vocative intonation patterns are sensitive to politeness factors</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2015">2015</year>
<volume>58</volume>
<fpage>68</fpage>
<lpage>83</lpage>
<pub-id pub-id-type="doi">10.1177/0023830914565441</pub-id>
</element-citation>
</ref>
<ref id="B11">
<label>11</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Brazil</surname>
<given-names>D.</given-names>
</name>
</person-group>
<source>The communicative value of intonation in English</source>
<year iso-8601-date="1985">1985</year>
<publisher-loc>Birmingham</publisher-loc>
<publisher-name>Bleakhouse Press and English Language Research</publisher-name>
</element-citation>
</ref>
<ref id="B12">
<label>12</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Bruce</surname>
<given-names>G.</given-names>
</name>
</person-group>
<source>Swedish word accents in sentence perspective</source>
<year iso-8601-date="1977">1977</year>
<publisher-loc>Lund</publisher-loc>
<publisher-name>Gleerup</publisher-name>
</element-citation>
</ref>
<ref id="B13">
<label>13</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>A.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Lee</surname>
<given-names>W.-S.</given-names>
</name>
<name>
<surname>Zee</surname>
<given-names>E.</given-names>
</name>
</person-group>
<article-title>What&#8217;s in a rise: Evidence for an off-ramp analysis of Dutch intonation</article-title>
<conf-name>Proceedings of the 17th International Congress of Phonetic Sciences [ICPhS XVII]</conf-name>
<year iso-8601-date="2011">2011</year>
<conf-loc>Hong Kong</conf-loc>
<conf-sponsor>Department of Chinese, Translation and Linguistics, City University of Hong Kong</conf-sponsor>
<fpage>448</fpage>
<lpage>451</lpage>
</element-citation>
</ref>
<ref id="B14">
<label>14</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>den Os</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>de Ruiter</surname>
<given-names>J.-P.</given-names>
</name>
</person-group>
<article-title>Pitch accent type matters for online processing of information status: Evidence from natural and synthetic speech</article-title>
<source>The Linguistic Review</source>
<year iso-8601-date="2007">2007</year>
<volume>24</volume>
<fpage>317</fpage>
<lpage>344</lpage>
<pub-id pub-id-type="doi">10.1515/TLR.2007.012</pub-id>
</element-citation>
</ref>
<ref id="B15">
<label>15</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chen</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Xu</surname>
<given-names>Y.</given-names>
</name>
</person-group>
<article-title>Production of weak elements in speech: Evidence from F0 patterns of neutral tone in Standard Chinese</article-title>
<source>Phonetica</source>
<year iso-8601-date="2006">2006</year>
<volume>63</volume>
<fpage>47</fpage>
<lpage>75</lpage>
<pub-id pub-id-type="doi">10.1159/000091406</pub-id>
</element-citation>
</ref>
<ref id="B16">
<label>16</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cole</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>New Methods for Prosodic Transcription: Capturing Variability as a Source of Information</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>8</elocation-id>
<fpage>1</fpage>
<lpage>29</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.29</pub-id>
</element-citation>
</ref>
<ref id="B17">
<label>17</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Collier</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>&#8217;t Hart</surname>
<given-names>H.</given-names>
</name>
</person-group>
<source>Cursus Nederlandse intonatie</source>
<year iso-8601-date="1980">1980</year>
<publisher-loc>Leuven</publisher-loc>
<publisher-name>Acco</publisher-name>
</element-citation>
</ref>
<ref id="B18">
<label>18</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Cruttenden</surname>
<given-names>A.</given-names>
</name>
</person-group>
<source>Intonation</source>
<year iso-8601-date="1997">1997</year>
<edition>2nd ed.</edition>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
</element-citation>
</ref>
<ref id="B19">
<label>19</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Crystal</surname>
<given-names>D.</given-names>
</name>
</person-group>
<source>Prosodic systems and intonation in English</source>
<year iso-8601-date="1969">1969</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9781139166973</pub-id>
</element-citation>
</ref>
<ref id="B20">
<label>20</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Dainora</surname>
<given-names>A.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Goldstein</surname>
<given-names>L. M.</given-names>
</name>
<name>
<surname>Whalen</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Best</surname>
<given-names>C. T.</given-names>
</name>
</person-group>
<chapter-title>Modeling intonation in English: A probabilistic approach to phonological competence</chapter-title>
<source>Papers in Laboratory Phonology 8: Varieties of Phonological Competence</source>
<year iso-8601-date="2006">2006</year>
<publisher-loc>Berlin/New York</publisher-loc>
<publisher-name>Mouton de Gruyter</publisher-name>
<fpage>107</fpage>
<lpage>132</lpage>
</element-citation>
</ref>
<ref id="B21">
<label>21</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dilley</surname>
<given-names>L. C.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R..</given-names>
</name>
<name>
<surname>Schepman</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Alignment of L and H in bitonal pitch accents: testing two hypotheses</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2005">2005</year>
<volume>33</volume>
<fpage>115</fpage>
<lpage>119</lpage>
<pub-id pub-id-type="doi">10.1016/j.wocn.2004.02.003</pub-id>
</element-citation>
</ref>
<ref id="B22">
<label>22</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Frota</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Surface and Structure: Transcribing Intonation within and across Languages</article-title>
<source>Laboratory Phonology: Journal of the Association for Laboratory Phonology</source>
<year iso-8601-date="2016">2016</year>
<volume>7</volume>
<issue>1</issue>
<elocation-id>7</elocation-id>
<fpage>1</fpage>
<lpage>19</lpage>
<pub-id pub-id-type="doi">10.5334/labphon.10</pub-id>
<comment>5</comment>
</element-citation>
</ref>
<ref id="B23">
<label>23</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gibbon</surname>
<given-names>D.</given-names>
</name>
</person-group>
<source>Perspectives on intonation analysis</source>
<year iso-8601-date="1976">1976</year>
<publisher-loc>Bern</publisher-loc>
<publisher-name>Lang</publisher-name>
</element-citation>
</ref>
<ref id="B24">
<label>24</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Goldsmith</surname>
<given-names>J. A.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Goyvaerts</surname>
<given-names>D.</given-names>
</name>
</person-group>
<chapter-title>English as a tone language</chapter-title>
<source>Phonology in the 80s</source>
<year iso-8601-date="1980">1980</year>
<publisher-loc>Ghent</publisher-loc>
<publisher-name>Story-Scientia</publisher-name>
<fpage>287</fpage>
<lpage>308</lpage>
<comment>(Unpublished mimeographed version, 1974)</comment>
</element-citation>
</ref>
<ref id="B25">
<label>25</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Grabe</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Post</surname>
<given-names>B.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Sampson</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>McCarthy</surname>
<given-names>D.</given-names>
</name>
</person-group>
<chapter-title>Intonational variation in the British Isles</chapter-title>
<source>Corpus linguistics: Readings in a widening discipline</source>
<year iso-8601-date="2004">2004</year>
<publisher-loc>London and New York</publisher-loc>
<publisher-name>Continuum International</publisher-name>
<fpage>474</fpage>
<lpage>481</lpage>
</element-citation>
</ref>
<ref id="B26">
<label>26</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Leading tones and downstep in English</article-title>
<source>Phonology</source>
<year iso-8601-date="1995">1995</year>
<volume>12</volume>
<fpage>183</fpage>
<lpage>233</lpage>
<pub-id pub-id-type="doi">10.1017/S0952675700002475</pub-id>
</element-citation>
</ref>
<ref id="B27">
<label>27</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Baumann</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Benzm&#252;ller</surname>
<given-names>R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>German intonation and autosegmental-metrical phonology</chapter-title>
<source>Prosodic typology: The phonology of intonation and phrasing</source>
<year iso-8601-date="2005">2005</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>55</fpage>
<lpage>83</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199249633.003.0003</pub-id>
</element-citation>
</ref>
<ref id="B28">
<label>28</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Grice</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Baumann</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Jagdfeld</surname>
<given-names>N.</given-names>
</name>
</person-group>
<article-title>Tonal association and derived nuclear accents: The case of downstepping contours in German</article-title>
<source>Lingua</source>
<year iso-8601-date="2009">2009</year>
<volume>119</volume>
<fpage>881</fpage>
<lpage>905</lpage>
<pub-id pub-id-type="doi">10.1016/j.lingua.2007.11.013</pub-id>
</element-citation>
</ref>
<ref id="B29">
<label>29</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<chapter-title>A semantic analysis of the nuclear tones of English</chapter-title>
<source>On the grammar and semantics of sentence accents</source>
<year iso-8601-date="1983">1983</year>
<publisher-loc>Dordrecht</publisher-loc>
<publisher-name>Foris</publisher-name>
<comment>Distributed by IULC. Included as ch. 5 in Gussenhoven, C. 1985</comment>
</element-citation>
</ref>
<ref id="B30">
<label>30</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Broe</surname>
<given-names>M. B.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
</person-group>
<chapter-title>The boundary tones are coming: On the non-peripheral realization of boundary tones</chapter-title>
<source>Papers in Laboratory Phonology V: Acquisition and the Lexicon</source>
<year iso-8601-date="2000">2000</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<fpage>132</fpage>
<lpage>151</lpage>
</element-citation>
</ref>
<ref id="B31">
<label>31</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<source>The phonology of tone and intonation</source>
<year iso-8601-date="2004">2004</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511616983</pub-id>
</element-citation>
</ref>
<ref id="B32">
<label>32</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<article-title>Semantic judgments as evidence for the intonational structure of Dutch</article-title>
<conf-name>Proceedings of Speech Prosody</conf-name>
<conf-date>2008</conf-date>
<year iso-8601-date="2008">2008</year>
<conf-loc>Campinas, Brazil</conf-loc>
<fpage>297</fpage>
<lpage>300</lpage>
</element-citation>
</ref>
<ref id="B33">
<label>33</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Rietveld</surname>
<given-names>A. C. M.</given-names>
</name>
</person-group>
<article-title>An experimental evaluation of two nuclear tone taxonomies</article-title>
<source>Linguistics</source>
<year iso-8601-date="1991">1991</year>
<volume>29</volume>
<fpage>423</fpage>
<lpage>449</lpage>
<pub-id pub-id-type="doi">10.1515/ling.1991.29.3.423</pub-id>
</element-citation>
</ref>
<ref id="B34">
<label>34</label>
<element-citation publication-type="webpage">
<person-group person-group-type="author">
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Rietveld</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Kerkhoff</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Terken</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>Transcription of Dutch intonation</article-title>
<year iso-8601-date="1999&#8211;2003">1999&#8211;2003</year>
<comment>Retrieved from <uri>todi.science.ru.nl</uri></comment>
</element-citation>
</ref>
<ref id="B35">
<label>35</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Hayes</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Lahiri</surname>
<given-names>A.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Sundberg</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Nord</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Carlson</surname>
<given-names>R.</given-names>
</name>
</person-group>
<chapter-title>Durationally specified intonation in English and Bengali</chapter-title>
<source>Music, language, speech, and brain</source>
<year iso-8601-date="1991">1991</year>
<publisher-loc>London</publisher-loc>
<publisher-name>Macmillan</publisher-name>
<fpage>78</fpage>
<lpage>91</lpage>
<pub-id pub-id-type="doi">10.1007/978-1-349-12670-5_7</pub-id>
</element-citation>
</ref>
<ref id="B36">
<label>36</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Kahn</surname>
<given-names>S. D.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<chapter-title>The intonational phonology of Bangladeshi Standard Bengali</chapter-title>
<source>Prosodic typology II: The phonology of intonation and phrasing</source>
<year iso-8601-date="2014">2014</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<fpage>81</fpage>
<lpage>117</lpage>
<pub-id pub-id-type="doi">10.1093/acprof:oso/9780199567300.003.0004</pub-id>
</element-citation>
</ref>
<ref id="B37">
<label>37</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<article-title>Stylized intonation</article-title>
<source>Language</source>
<year iso-8601-date="1978">1978</year>
<volume>54</volume>
<fpage>517</fpage>
<lpage>540</lpage>
<pub-id pub-id-type="doi">10.1353/lan.1978.0056</pub-id>
</element-citation>
</ref>
<ref id="B38">
<label>38</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<source>The structure of intonational meaning: Evidence from English</source>
<year iso-8601-date="1980">1980</year>
<publisher-loc>Bloomington</publisher-loc>
<publisher-name>Indiana University Press</publisher-name>
</element-citation>
</ref>
<ref id="B39">
<label>39</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<article-title>Phonological features of intonational peaks</article-title>
<source>Language</source>
<year iso-8601-date="1983">1983</year>
<volume>59</volume>
<fpage>721</fpage>
<lpage>759</lpage>
<pub-id pub-id-type="doi">10.2307/413371</pub-id>
</element-citation>
</ref>
<ref id="B40">
<label>40</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<source>Intonational phonology</source>
<year iso-8601-date="2008">2008</year>
<edition>2nd ed.</edition>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511808814</pub-id>
</element-citation>
</ref>
<ref id="B41">
<label>41</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
<name>
<surname>Schepman</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>&#8220;Sagging transitions&#8221; between high pitch accents in English: Experimental evidence</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2003">2003</year>
<volume>31</volume>
<fpage>81</fpage>
<lpage>112</lpage>
<pub-id pub-id-type="doi">10.1016/S0095-4470(02)00073-6</pub-id>
</element-citation>
</ref>
<ref id="B42">
<label>42</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Leben</surname>
<given-names>W. R.</given-names>
</name>
</person-group>
<article-title>The tones of English intonation</article-title>
<source>Linguistic Analysis</source>
<year iso-8601-date="1976">1976</year>
<volume>2</volume>
<fpage>69</fpage>
<lpage>107</lpage>
</element-citation>
</ref>
<ref id="B43">
<label>43</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>Liberman</surname>
<given-names>M. Y.</given-names>
</name>
</person-group>
<source>The intonational system of English, MIT dissertation</source>
<year iso-8601-date="1975">1975</year>
<publisher-name>Garland Publishing</publisher-name>
<comment>Published 1979 by</comment>
</element-citation>
</ref>
<ref id="B44">
<label>44</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lickley</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Schepman</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<article-title>Alignment of &#8216;phrase accent&#8217; low in Dutch falling rising questions: Theoretical and methodological implications</article-title>
<source>Language and Speech</source>
<year iso-8601-date="2005">2005</year>
<volume>48</volume>
<fpage>157</fpage>
<lpage>183</lpage>
<pub-id pub-id-type="doi">10.1177/00238309050480020201</pub-id>
</element-citation>
</ref>
<ref id="B45">
<label>45</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Mayo</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Aylett</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Botinis</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Kouroupetroglou</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Carayannis</surname>
<given-names>G.</given-names>
</name>
</person-group>
<chapter-title>Prosodic transcription of Glasgow English: An evaluation study of GlaToBI</chapter-title>
<source>Intonation: Theory, models and applications. Proceedings of an ESCA Workshop</source>
<year iso-8601-date="1997">1997</year>
<publisher-loc>Athens</publisher-loc>
<publisher-name>ESCA and University of Athens, Department of Informatics</publisher-name>
<fpage>231</fpage>
<lpage>234</lpage>
</element-citation>
</ref>
<ref id="B46">
<label>46</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>McCarthy</surname>
<given-names>J. J.</given-names>
</name>
<name>
<surname>Prince</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Generalized alignment</article-title>
<source>Department of Linguistics Faculty Publication Series</source>
<year iso-8601-date="1993">1993</year>
<pub-id pub-id-type="doi">10.1007/978-94-017-3712-8_4</pub-id>
<comment>Paper 12</comment>
</element-citation>
</ref>
<ref id="B47">
<label>47</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Nolan</surname>
<given-names>F.</given-names>
</name>
<name>
<surname>Grabe</surname>
<given-names>E.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Botinis</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Kouroupetroglou</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Carayannis</surname>
<given-names>G.</given-names>
</name>
</person-group>
<chapter-title>Can ToBI transcribe intonational variation in English?</chapter-title>
<source>Intonation: Theory, models and applications. Proceedings of an ESCA Workshop</source>
<year iso-8601-date="1997">1997</year>
<publisher-loc>Athens</publisher-loc>
<publisher-name>ESCA and University of Athens, Department of Informatics</publisher-name>
<fpage>259</fpage>
<lpage>262</lpage>
</element-citation>
</ref>
<ref id="B48">
<label>48</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>O&#8217;Connor</surname>
<given-names>J. D.</given-names>
</name>
<name>
<surname>Arnold</surname>
<given-names>G. F.</given-names>
</name>
</person-group>
<source>Intonation of colloquial English</source>
<year iso-8601-date="1973">1973</year>
<edition>(2nd ed.)</edition>
<publisher-loc>London</publisher-loc>
<publisher-name>Longman</publisher-name>
</element-citation>
</ref>
<ref id="B49">
<label>49</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Palmer</surname>
<given-names>H. E.</given-names>
</name>
</person-group>
<source>English intonation with systematic exercises</source>
<year iso-8601-date="1922">1922</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Heffer</publisher-name>
</element-citation>
</ref>
<ref id="B50">
<label>50</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Peters</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Hanssen</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<article-title>The timing of nuclear falls: Evidence from Dutch, West Frisian, Dutch Low Saxon, German Low Saxon, and High German</article-title>
<source>Laboratory Phonology</source>
<year iso-8601-date="2015">2015</year>
<volume>6</volume>
<fpage>1</fpage>
<lpage>52</lpage>
<pub-id pub-id-type="doi">10.1515/lp-2015-0004</pub-id>
</element-citation>
</ref>
<ref id="B51">
<label>51</label>
<element-citation publication-type="thesis">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
</person-group>
<source>The phonetics and phonology of English intonation (Unpublished doctoral dissertation)</source>
<year iso-8601-date="1980">1980</year>
<publisher-name>MIT</publisher-name>
</element-citation>
</ref>
<ref id="B52">
<label>52</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
</person-group>
<article-title>Alignment and prosodic heads</article-title>
<conf-name>Proceedings of the Eastern States Conference on Formal Linguistics (ESCOL)</conf-name>
<year iso-8601-date="1993">1993</year>
<conf-sponsor>Linguistics Graduate Student Association, Cornell</conf-sponsor>
<fpage>268</fpage>
<lpage>286</lpage>
</element-citation>
</ref>
<ref id="B53">
<label>53</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Horne</surname>
<given-names>M.</given-names>
</name>
</person-group>
<chapter-title>Tonal elements and their alignment</chapter-title>
<source>Intonation: Theory and experiment</source>
<year iso-8601-date="2000">2000</year>
<publisher-loc>Dordrecht</publisher-loc>
<publisher-name>Kluwer</publisher-name>
<fpage>11</fpage>
<lpage>36</lpage>
<pub-id pub-id-type="doi">10.1007/978-94-015-9413-4_2</pub-id>
</element-citation>
</ref>
<ref id="B54">
<label>54</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Pitrelli</surname>
<given-names>J. F.</given-names>
</name>
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>Evaluation of prosodic transcription labeling reliability in the ToBI framework</article-title>
<conf-name>Proceedings of the International Conference on Spoken Language Processing (ICSLP)</conf-name>
<year iso-8601-date="1994">1994</year>
<conf-loc>Yokohama, Japan</conf-loc>
<fpage>123</fpage>
<lpage>126</lpage>
</element-citation>
</ref>
<ref id="B55">
<label>55</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Riad</surname>
<given-names>T.</given-names>
</name>
</person-group>
<source>Swedish phonology</source>
<year iso-8601-date="2014">2014</year>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
</element-citation>
</ref>
<ref id="B56">
<label>56</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Ritchart</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Arvaniti</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>The form and use of uptalk in Southern California English</article-title>
<conf-name>Proceedings of Speech Prosody 7</conf-name>
<year iso-8601-date="2014">2014</year>
<conf-loc>Dublin</conf-loc>
<comment>Retrieved from <uri>http://www.speechprosody2014.org</uri></comment>
</element-citation>
</ref>
<ref id="B57">
<label>57</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Shattuck-Hufnagel</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>Dilley</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Veilleux</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Brugos</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Speer</surname>
<given-names>R.</given-names>
</name>
</person-group>
<article-title>F0 peaks and valleys aligned with non-prominent syllables can influence perceived prominence in adjacent syllables</article-title>
<conf-name>Proceedings of Speech Prosody</conf-name>
<year iso-8601-date="2004">2004</year>
<volume>2</volume>
<fpage>705</fpage>
<lpage>708</lpage>
</element-citation>
</ref>
<ref id="B58">
<label>58</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Silverman</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Beckman</surname>
<given-names>M. E.</given-names>
</name>
<name>
<surname>Pitrelli</surname>
<given-names>J. F.</given-names>
</name>
<name>
<surname>Ostendorf</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Wightman</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Price</surname>
<given-names>P. J.</given-names>
</name>
<name>
<surname>Pierrehumbert</surname>
<given-names>J. B.</given-names>
</name>
<name>
<surname>Hirschberg</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>TOBI: A standard for labelling English prosody</article-title>
<conf-name>Proceedings of ICSLP</conf-name>
<year iso-8601-date="1992">1992</year>
<conf-loc>Banff, Alberta, Canada</conf-loc>
</element-citation>
</ref>
<ref id="B59">
<label>59</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Steedman</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Structure and intonation</article-title>
<source>Language</source>
<year iso-8601-date="1991">1991</year>
<volume>67</volume>
<fpage>260</fpage>
<lpage>296</lpage>
<pub-id pub-id-type="doi">10.1353/lan.1991.0098</pub-id>
</element-citation>
</ref>
<ref id="B60">
<label>60</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Steedman</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>The surface-compositional semantics of English intonation</article-title>
<source>Language</source>
<year iso-8601-date="2014">2014</year>
<volume>90</volume>
<fpage>2</fpage>
<lpage>57</lpage>
<pub-id pub-id-type="doi">10.1353/lan.2014.0010</pub-id>
</element-citation>
</ref>
<ref id="B61">
<label>61</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>&#8217;t Hart</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Collier</surname>
<given-names>R.</given-names>
</name>
</person-group>
<source>Cursus Nederlandse Intonatie</source>
<year iso-8601-date="1980">1980</year>
<publisher-loc>Louvain</publisher-loc>
<publisher-name>Acco</publisher-name>
</element-citation>
</ref>
<ref id="B62">
<label>62</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>&#8217;t Hart</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Collier</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Cohen</surname>
<given-names>A.</given-names>
</name>
</person-group>
<source>A perceptual study of intonation: An experimental-phonetic approach to speech melody</source>
<year iso-8601-date="1990">1990</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<pub-id pub-id-type="doi">10.1017/CBO9780511627743</pub-id>
</element-citation>
</ref>
<ref id="B63">
<label>63</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Tench</surname>
<given-names>P.</given-names>
</name>
</person-group>
<source>The intonation systems of English</source>
<year iso-8601-date="1996">1996</year>
<publisher-loc>London/New York</publisher-loc>
<publisher-name>Cassell</publisher-name>
</element-citation>
</ref>
<ref id="B64">
<label>64</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Trim</surname>
<given-names>J. L. M.</given-names>
</name>
</person-group>
<article-title>Major and minor tone groups in English</article-title>
<source>Le Maitre Phon&#233;tique</source>
<year iso-8601-date="1959">1959</year>
<volume>112</volume>
<fpage>26</fpage>
<lpage>29</lpage>
</element-citation>
</ref>
<ref id="B65">
<label>65</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Uldall</surname>
<given-names>E.</given-names>
</name>
</person-group>
<chapter-title>Review of R. Kingdon</chapter-title>
<source>The groundwork of English intonation</source>
<year iso-8601-date="1961">1961</year>
<volume>13</volume>
<publisher-name>Oxford University Press</publisher-name>
<fpage>214</fpage>
<lpage>218</lpage>
<comment>1985. Archivum Linguisticum</comment>
</element-citation>
</ref>
<ref id="B66">
<label>66</label>
<element-citation publication-type="confproc">
<person-group person-group-type="author">
<name>
<surname>Vanderslice</surname>
<given-names>R.</given-names>
</name>
</person-group>
<person-group person-group-type="editor">
<name>
<surname>Rigault</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Charboneau</surname>
<given-names>R.</given-names>
</name>
</person-group>
<article-title>The binary suprasegmental features of English</article-title>
<conf-name>Proceedings of the Seventh International Congress of Phonetic Sciences</conf-name>
<year iso-8601-date="1972">1972</year>
<conf-loc>The Hague/Paris</conf-loc>
<conf-sponsor>Mouton</conf-sponsor>
<fpage>1052</fpage>
<lpage>1057</lpage>
<comment>(Montreal 1971)</comment>
</element-citation>
</ref>
<ref id="B67">
<label>67</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Vanderslice</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Pierson</surname>
<given-names>L. S.</given-names>
</name>
</person-group>
<article-title>Prosodic features of Hawaiian</article-title>
<source>English. Quarterly Journal of Speech</source>
<year iso-8601-date="1967">1967</year>
<volume>53</volume>
<fpage>156</fpage>
<lpage>166</lpage>
<pub-id pub-id-type="doi">10.1080/00335636709382828</pub-id>
</element-citation>
</ref>
<ref id="B68">
<label>68</label>
<element-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>van de Ven</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Gussenhoven</surname>
<given-names>C.</given-names>
</name>
</person-group>
<article-title>The timing of the final rise in falling-rising intonation contours in Dutch</article-title>
<source>Journal of Phonetics</source>
<year iso-8601-date="2011">2011</year>
<volume>39</volume>
<fpage>225</fpage>
<lpage>236</lpage>
<pub-id pub-id-type="doi">10.1016/j.wocn.2011.01.006</pub-id>
</element-citation>
</ref>
<ref id="B69">
<label>69</label>
<element-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Wells</surname>
<given-names>J. C.</given-names>
</name>
</person-group>
<source>English intonation: An introduction</source>
<year iso-8601-date="2006">2006</year>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
</element-citation>
</ref>
</ref-list>
</back>
</article>