<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.2 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.2/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.2" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology: Journal of the Association for Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Open Library of Humanities</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.16995/labphon.6463</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Gridlines approach for dynamic analysis in speech ultrasound data: A multimodal app</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author">
<name>
<surname>Gonzalez</surname>
<given-names>Simon</given-names>
</name>
<email>simon.gonzalez@anu.edu.au</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>The Australian National University, AU</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2021-11-19">
<day>19</day>
<month>11</month>
<year>2021</year>
</pub-date>
<pub-date pub-type="collection">
<year>2021</year>
</pub-date>
<volume>12</volume>
<issue>1</issue>
<elocation-id>16</elocation-id>
<history>
<date date-type="received" iso-8601-date="2019-09-30">
<day>30</day>
<month>09</month>
<year>2019</year>
</date>
<date date-type="accepted" iso-8601-date="2021-06-03">
<day>03</day>
<month>06</month>
<year>2021</year>
</date>
</history>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2021 The Author(s)</copyright-statement>
<copyright-year>2021</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.16995/labphon.6463/"/>
<abstract>
<p>Having access to efficient technologies is essential for the accurate description and analysis of articulatory speech patterns. In the area of tongue ultrasound studies, the visualization/analysis processes generally require a solid knowledge of programming languages as well as a deep understanding of articulatory phenomena. This demands the use of a variety of programs for an efficient use of the data collected. In this paper I introduce a multimodal app for visualizing and analyzing tongue contours: UVA&#8212;Ultrasound Visualization and Analysis. This app combines the computational power of R and the interactivity of Shiny web apps to allow users to manipulate and explore tongue ultrasound data using cutting-edge methods. One of the greatest strengths of the app is that it has the capability of being modified to adapt to the users&#8217; needs. This has potential as an innovative tool for diverse academic and industry audiences.</p>
</abstract>
<kwd-group>
<kwd>Tongue ultrasound</kwd>
<kwd>gridlines</kwd>
<kwd>shiny app</kwd>
<kwd>dynamic analysis</kwd>
<kwd>distance</kwd>
<kwd>displacement</kwd>
<kwd>velocity</kwd>
<kwd>acceleration</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec>
<title>1. Introduction</title>
<p>Speech analysis technologies have played a fundamental role in the understanding of how human language uses all available articulators for communication. Within the literature, there has been a strong emphasis on studying the tongue as one of the most important organs in speech production. There is a range of techniques which can be used for this purpose, including X-ray (<xref ref-type="bibr" rid="B10">Bressmann, Koch, Ratner, Seigel, &amp; Binkofski, 2015</xref>; <xref ref-type="bibr" rid="B66">Verma, Tandon, Agrawal, &amp; Prabhat, 2012</xref>), electropalatography (<xref ref-type="bibr" rid="B7">Barberena et al., 2017</xref>; <xref ref-type="bibr" rid="B19">Gibbon, Lee, &amp; Yuen, 2010</xref>; <xref ref-type="bibr" rid="B43">N. R. Miller, Reyes-Aldasoro, &amp; Verhoeven, 2019</xref>; <xref ref-type="bibr" rid="B65">Verhoeven, Miller, Daems, &amp; Reyes-Aldasoro, 2019</xref>), magnetic resonance imaging (<xref ref-type="bibr" rid="B23">Hewer, Wuhrer, Steiner, &amp; Richmond, 2018</xref>; <xref ref-type="bibr" rid="B36">Lim et al., 2019</xref>; <xref ref-type="bibr" rid="B38">Maekawa, 2019</xref>; <xref ref-type="bibr" rid="B51">Proctor, Lo, &amp; Narayanan, 2015</xref>), electromagnetic articulography (<xref ref-type="bibr" rid="B27">Katz, Mehta, &amp; Wood, 2017</xref>; <xref ref-type="bibr" rid="B29">Kocharov &amp; Evdokimova, 2019</xref>; <xref ref-type="bibr" rid="B43">N. R. Miller et al., 2019</xref>; <xref ref-type="bibr" rid="B56">Shadle, Proctor, &amp; Iskarous, 2008</xref>; <xref ref-type="bibr" rid="B72">Zeroual, Hoole, &amp; Gafos, 2019</xref>), and ultrasound (<xref ref-type="bibr" rid="B17">Diskin et al., 2019</xref>; <xref ref-type="bibr" rid="B41">Mielke, Carignan, &amp; Thomas, 2017</xref>; <xref ref-type="bibr" rid="B74">Zharkova, Gibbon, &amp; Lee, 2017</xref>). Among these, ultrasound is used to examine correlates between tongue articulations and acoustic or phonological phenomena. Gick, Campbell, and Oh (<xref ref-type="bibr" rid="B20">2001</xref>) note that phonological contrast may not always be observable in the acoustic signal but may be observable in other levels, like articulatory gestures. Ultrasound offers the opportunity to observe and analyze these gestures and also allows for the observation and analysis of tongue articulations in detail. An advantage of ultrasound over other tongue imaging techniques (e.g., electropalatography, electromagnetic midsagittal articulography, and X-ray microbeam) is that a great length of the tongue contour can be imaged (<xref ref-type="bibr" rid="B14">Davidson, 2006</xref>). It is widely used to measure an extensive mid-sagittal section of the tongue, allowing measurements of the tongue surface from most anterior to most posterior sections by examining upward and downward movements. The only limitation of ultrasound imaging is mainly on the most anterior portions of the tongue, such as maximum constriction on dental segments and in some cases, advanced alveolar realizations (<xref ref-type="bibr" rid="B21">Gonzalez, 2015</xref>). Mid-sagittal sections allow measuring three main parts of the tongue, namely, the front part of the tongue, the body, and the dorsal section.</p>
<p>Ultrasound tongue imaging has traditionally been analyzed using either static or dynamic approaches. One of the main techniques in static approaches is the comparison of tongue contours at specific articulatory landmarks (<xref ref-type="bibr" rid="B3">Alwabari, 2019</xref>; <xref ref-type="bibr" rid="B43">N. R. Miller et al., 2019</xref>; <xref ref-type="bibr" rid="B46">Oakley, 2019</xref>; <xref ref-type="bibr" rid="B55">Roon &amp; Whalen, 2019</xref>). An advantage of this approach is that it allows for the comparison of gestural differences at a specific time of the speech process previously established in the study, for example, differences between vowels (<xref ref-type="bibr" rid="B16">Decker &amp; Nycz, 2012</xref>; <xref ref-type="bibr" rid="B41">Mielke et al., 2017</xref>), consonants (<xref ref-type="bibr" rid="B1">Ahn, 2015</xref>, <xref ref-type="bibr" rid="B2">2018</xref>; <xref ref-type="bibr" rid="B54">Recasens &amp; Rodri&#769;guez, 2019</xref>), or phonological contexts (<xref ref-type="bibr" rid="B14">Davidson, 2006</xref>; <xref ref-type="bibr" rid="B17">Diskin et al., 2019</xref>). On the other hand, depending on the research questions to be addressed, dynamic approaches can capture time-series characteristics which can be crucial for identifying distinctions between segments that in a static approach may not be observed. Dynamic approaches can therefore be used to study the kinetics of speech sounds (<xref ref-type="bibr" rid="B30">Kochetov, Faytak, &amp; Nara, 2019</xref>; <xref ref-type="bibr" rid="B34">S. R. Li et al., 2019</xref>; <xref ref-type="bibr" rid="B42">A. Miller &amp; Finch, 2011</xref>). One main advantage of this approach is that it can measure articulatory data in an articulatory continuum. Data for this approach generally require more preparation than in static approaches, which makes it computationally more expensive and time consuming, depending on the workflow chosen.</p>
<sec>
<title>1.1. The general workflow of ultrasound studies</title>
<p>The workflow within an ultrasound study varies depending on many factors, mainly of the research questions and the resources available for the study. However, the process of ultrasound analysis can generally be broken down into five main stages: data collection, contour extraction,<xref ref-type="fn" rid="n1">1</xref> data wrangling, visualization, and analysis.<xref ref-type="fn" rid="n2">2</xref> At every stage, there are important challenges for any researcher, and here I comment on key challenges that the present study aims to address. It is important to note that these stages are not necessarily sequential, since in some cases they can happen simultaneously. During data collection, one important question is the quality of the data and the frame rate at which contours are imaged. A low frame rate poses the challenge of missing key articulatory landmarks, depending on the phonological phenomena analyzed. For example, for trills, a high frame rate is required to capture more accurate gestural timing, as in Proctor (<xref ref-type="bibr" rid="B50">2009</xref>). However, for other phonological phenomena, such as vowels, a lower frame rate can be used. The challenge of high frame rates is that they can be computationally too expensive, and not knowing the optimal rate in advance can result in frame rates that are higher than necessary (meaning that the resource cost exceeds the benefit of having a higher frame rate). The second stage is contour extraction. Researchers make use of many available computer programs such as EdgeTrack (<xref ref-type="bibr" rid="B33">M. Li, Kambhamettu, &amp; Stone, 2005</xref>) and the Articulate Assistant Advanced (AAA) software (<xref ref-type="bibr" rid="B70">Wrench, 2012</xref>), which are used for semi-automatic extraction of tongue contours. For an expanded review on different methods and studies and a new approach for fully automated extractions, see Karimi, Menard, and Laporte (<xref ref-type="bibr" rid="B26">2019</xref>). These programs are not 100% accurate and a manual correction stage is part of the process. The third stage is data wrangling, when the data is prepared for analysis. Among the programming languages used for wrangling, R is widely used (cf. <xref ref-type="bibr" rid="B14">Davidson, 2006</xref>; <xref ref-type="bibr" rid="B16">Decker &amp; Nycz, 2012</xref>; <xref ref-type="bibr" rid="B37">Lin, Beddor, &amp; Coetzee, 2014</xref>; <xref ref-type="bibr" rid="B49">Pini, Spreafico, Vantini, &amp; Vietti, 2019</xref>; <xref ref-type="bibr" rid="B54">Recasens &amp; Rodri&#769;guez, 2019</xref>). Developing tools using this programming language therefore can be beneficial to the research community.</p>
<p>The fourth stage is visualization. The importance of this stage is that it is strongly connected to the analysis to be carried out in further stages. This means that the visualization is crucial not only to identify patterns in tongue kinematics, but also to establish the most appropriate analysis approach and methods. Three key aspects are strongly considered in the visualization stage.</p>
<sec>
<title>1.1.1. Field of View</title>
<p>Field of View is an important parameter for ultrasound analysis. It defines the extent of the comparable contours observed on the graphs window at any given moment. It is measured as a horizontal angle (angle aperture), which corresponds with the virtual image origin of the ultrasound probe. The field of view depends on the articulatory phenomena examined. For example, in the case of coronal segments, it is preferable that the field capture activity towards the mid and front sections of the tongue. In the case of velar segments, it is preferable that the field be more retracted to capture more activity towards the back of the tongue. Establishing the field of view is crucial because if it is not properly established, there may be key articulations missed in the range chosen for observation. It is important to highlight the fact that the field of view must be considered before purchasing a machine and transducer since this cannot always be changed.</p>
<p>In this paper, I have implemented a functionality for users to specify the range or subsection of the original Field of View. To distinguish the internal capability of the app from the field of view, I refer to this as the Analysis Fan View (hereafter AFV). This refers to the sections from the original tongue contours that the user chooses to focus on. For this purpose, the angle origin in the AFV is created based on the tongue contours uploaded in the app, which does not directly correspond with the angle origin from the surface of the probe used in the data collection. This is developed in more detail in Section 3.3.2.</p>
</sec>
<sec>
<title>1.1.2. Landmarks definition</title>
<p>Articulatory landmarks are relevant because they can be used to define the areas in the tongue to be analyzed. Tongue ultrasound imaging poses challenges in relation to defining articulatory landmarks. Since the tongue moves as a whole unit and there are no hard-defined sections, as there are in passive articulators (e.g., teeth, alveolar ridge, soft palate), landmark definitions are necessary to accurately interpret the dynamics of the tongue (c.f. <xref ref-type="bibr" rid="B28">Kier &amp; Smith, 1985</xref>; <xref ref-type="bibr" rid="B61">Stone &amp; Murano, 2007</xref>). In ultrasound research, landmarks can be acoustic and/or articulatory. In the case of acoustic landmarks (cf. <xref ref-type="bibr" rid="B15">Dawson, Tiede, &amp; Whalen, 2016</xref>; <xref ref-type="bibr" rid="B32">Lawson &amp; Stuart-Smith, 2019</xref>; <xref ref-type="bibr" rid="B39">Mark&#243;, Bart&#243;k, Csap&#243;, Deme, &amp; Gr&#225;czi, 2019</xref>; <xref ref-type="bibr" rid="B44">Mizoguchi, Tiede, &amp; Whalen, 2019</xref>; <xref ref-type="bibr" rid="B73">Zharkova, 2013</xref>), acoustic cues are the bases for defining moments during the articulatory process, for example, the acoustic time of the mid-point of a vowel (<xref ref-type="bibr" rid="B41">Mielke et al., 2017</xref>) or the maximum constriction of a consonant located at the mid-point of its duration. On the other hand, articulatory landmarks are defined by the gestural behaviour of the segment. For example, the constriction location in low front vowels is located at the still frame showing the most retracted tongue contour on the tongue dorsum (as in <xref ref-type="bibr" rid="B16">Decker &amp; Nycz, 2012</xref>), and the maximum constriction of a velar stop is located at the frame showing the most raised tongue body during stop closure (as in <xref ref-type="bibr" rid="B14">Davidson, 2006</xref>). These two approaches for landmark definition are not mutually exclusive. It is common practice in the field to use mixed approaches in which both acoustic and articulatory cues are considered.</p>
</sec>
<sec>
<title>1.1.3. Articulatory landmarks for analysis</title>
<p>The articulatory moment(s), or the time frame, is another important parameter and it refers to the dynamic windows for analysis, for example, analyzing transitions between a vowel and the maximum constriction of the following consonant. In static approaches, the time window is just one screenshot at a given moment, e.g., maximum constriction of a consonant or the onset of a vowel. The definition of these articulatory landmarks strongly depends on the phonological phenomena observed and the questions to be answered. In the case of dynamic studies, the purpose is to measure gestural patterns from point A to point B of a given sequence.</p>
<p>The last stage in the ultrasound workflow is the analysis stage. This includes the type of measurement and the statistical approaches for analysis&#8212;qualitative, quantitative, or a combination of both. For example, in static approaches, studies have used Principal Component Analysis (<xref ref-type="bibr" rid="B60">Stone, Goldstein, &amp; Zhang, 1997</xref>) and Smoothing Splines to compare tongue contours and measure differences between segments (<xref ref-type="bibr" rid="B14">Davidson, 2006</xref>; <xref ref-type="bibr" rid="B16">Decker &amp; Nycz, 2012</xref>), and dynamic studies have looked at tongue displacement (<xref ref-type="bibr" rid="B21">Gonzalez, 2015</xref>) and velocities (<xref ref-type="bibr" rid="B62">Strycharczuk &amp; Scobbie, 2015</xref>).</p>
</sec>
</sec>
</sec>
<sec>
<title>2. Motivation and main purpose</title>
<p>As observed above, there are many factors to be considered for a solid analysis of ultrasound data in speech research. This requires strong skills both in linguistic phenomena and programming analytical tools. Data analysis is achieved by using a combination of different software and scripts built in various programming languages. It is important to note that there are standalone programs which can efficiently do ultrasound analysis and incorporate different stages simultaneously, for example the Articulate Assistant Advanced (AAA) software (<xref ref-type="bibr" rid="B70">Wrench, 2012</xref>), which is powerful and of great use. However, these programs generally require specialized equipment and they do not tend to be available for code expansion/modification based on users&#8217; needs. Two relevant stand-alone <italic>R</italic> libraries have been developed to facilitate data processing from AAA. These are rticulate: Ultrasound Tongue Imaging in R (<xref ref-type="bibr" rid="B13">Coretta, 2020</xref>) and ultRa (<xref ref-type="bibr" rid="B8">Beare, 2018</xref>), which offer a range of functions to import and manipulate ultrasound data. In this paper, I introduce UVA: Ultrasound Visualization and Analysis, which can be accessed and downloaded from <ext-link ext-link-type="uri" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://github.com/simongonzalez/uva">https://github.com/simongonzalez/uva</ext-link>. The app implements a wide range of visualization parameters and carries out analysis of speech ultrasound data for both static and dynamic approaches. The aim is to present users with full control of visualization and analysis parameters to carry out efficient analysis of their data. It is a standalone environment which can be freely accessible and released as open source. This allows for any further fine-tuning of the code as well as implementation of new methods by modifying the code. This app is targeted at phoneticians analyzing static and dynamic data of tongue ultrasound images as well as speech practitioners interested in measuring tongue images for clinical purposes. In relation to the different stages of the workflow explained above, the app aims to enhance and facilitate mainly stages four (Visualization) and five (Analysis) with a tagging functionality for stage three (Data Wrangling). Since it does not offer collection and contour extraction, users are assumed to have tongue contours extracted in <italic>xy</italic> Cartesian coordinates, either in pixels and/or millimetres. In Sections 3 and 4, I describe in detail the app architecture and present a small sample study.</p>
</sec>
<sec>
<title>3. App architecture</title>
<p>The app implements both static and dynamic approaches from Gonzalez (<xref ref-type="bibr" rid="B21">2015</xref>) in a single program. The code used in the app was developed using open-source tools, combining the analytical capacities of the <italic>R</italic> environment (<xref ref-type="bibr" rid="B52">R Core Team, 2018</xref>) and the web-app capabilities of the Shiny library (<xref ref-type="bibr" rid="B11">Chang, Cheng, Allaire, Xie, &amp; McPherson, 2019</xref>), which is used for the creation of java-based web applications. The motivation is that R is a programming language widely used for speech analysis and Shiny allows the creation of apps that can have both the cutting-edge interactive interfaces and the power of the programming language analysis. The strength of the Shiny application framework is that it is intrinsically reactive in its programming. This means that it links input and output data and updates to the outputs (figures and tables) without refreshing the program or uploading new data, hence <italic>reacting</italic> to users&#8217; actions on the interface. All the programming actions in the background are translated as visual and graphic outputs at the front end, such as click-on buttons, sliders, drop-down menus. This allows users to explore data in more efficient and sophisticated ways without requiring ample knowledge of R programming. Users are only required to have basic knowledge of R such as opening and running apps.</p>
<p>For the proper working of the app, the input data must meet specific requirements. First of all, the number of individual points for each contour must start with 1 and end in 100. This is relevant specially for AAA users in which tongue contours do not always start with 1 but can start in higher numbers depending on the gridline capturing that point. Secondly, the app has been optimized to work with cartesian data. For future stages, I aim to implement compatibility with polar coordinates (cf. <xref ref-type="bibr" rid="B24">Heyne &amp; Derrick, 2015</xref>; <xref ref-type="bibr" rid="B40">Mielke, 2015</xref>).</p>
<p>In terms of the output data, the app automatically exports wrangled tongue contours with corresponding gridlines in a working folder labelled workingFiles, located within the main directory. All plots can be downloaded from the app in six formats: bmp, jpeg, pdf, png, svg, and tiff. For these images, users can change the width and height of the files.</p>
<p>The app requires a specific structure of the data, which can be structured in four levels as shown in Figure <xref ref-type="fig" rid="F1">1</xref>.</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Structure levels for the input data.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g1.png"/>
</fig>
<p>First, we have the Speaker level. It is a requirement to have at least one speaker in the data, this is, the data cannot have frames or sequences without being assigned to a speaker. In this sense, the speaker level is the overarching class in the data. The second level is the Segment. The data needs at least two segments to compare. It can be either two consonants or two vowels, or one segment in different conditions, for example, consonant /t/ in final and non-final position. The app requires each of these segments to have at least two <italic>repetitions</italic> to be analyzed, which is the third level. Segments may or may not have the same number of repetitions. For example, segment A has three repetitions and segment B has four repetitions. The app accounts for this difference using intrinsic and extrinsic criteria (See Figure <xref ref-type="fig" rid="F2">2</xref>). If it is intrinsic, only the same number of repetitions are taken into account. As in the example, only three repetitions per segment are analyzed, and the fourth one for segment B is ignored. On the other hand, if the extrinsic is selected, then all repetitions are analyzed irrespective of their uneven number of tokens.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Intrinsic and Extrinsic criteria for repetitions.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g2.png"/>
</fig>
<p>The fourth level is the number of frames. This refers to the number of frames per repetition for each segment. At least one frame per repetition is needed for the program to work properly. Similar to the number of repetitions, frames also follow an intrinsic and extrinsic selection. For each frame, the app reads four columns from the input data (See Table <xref ref-type="table" rid="T1">1</xref>). The first one is the number of the frame (&#8220;frame&#8221; column). The second one is the Cartesian Coordinate to which the measurement value in column &#8220;mm&#8221; is assigned. The &#8220;coord&#8221; column specifies whether it is the <italic>x</italic> or <italic>y</italic> coordinate. The third is the &#8220;point&#8221; column and it specifies the single point in the tongue contour. In the case of contours extracted from EdgeTrak, each trace had 100 points in the data tested. The fourth column read to create the frame is the &#8220;mm&#8221; column. This column stores the <italic>x</italic> or <italic>y</italic> value of the point in the contour. For columns repetition, frame, and point, all counts must start with 1 without any skipping. For instance, the program will not work properly if one segment has repetitions 1 and 3 (repetition 2 is missing), or repetitions 2 and 3 (it does not start with repetition 1).</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>Sample data format for the input data showing the first five observations.</p>
</caption>
<table>
<tr>
<th align="left" valign="top">speaker</th>
<th align="left" valign="top">segment</th>
<th align="center" valign="top">repetition</th>
<th align="center" valign="top">frame</th>
<th align="center" valign="top">coord</th>
<th align="center" valign="top">point</th>
<th align="center" valign="top">mm</th>
</tr>
<tr>
<td colspan="7"><hr/></td>
</tr>
<tr>
<td align="left" valign="top">1</td>
<td align="left" valign="top">s</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">x</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">48.15356</td>
</tr>
<tr>
<td align="left" valign="top">1</td>
<td align="left" valign="top">s</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">x</td>
<td align="right" valign="top">2</td>
<td align="right" valign="top">48.68272</td>
</tr>
<tr>
<td align="left" valign="top">1</td>
<td align="left" valign="top">s</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">x</td>
<td align="right" valign="top">3</td>
<td align="right" valign="top">49.47646</td>
</tr>
<tr>
<td align="left" valign="top">1</td>
<td align="left" valign="top">s</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">x</td>
<td align="right" valign="top">4</td>
<td align="right" valign="top">50.00562</td>
</tr>
<tr>
<td align="left" valign="top">1</td>
<td align="left" valign="top">s</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">1</td>
<td align="right" valign="top">x</td>
<td align="right" valign="top">5</td>
<td align="right" valign="top">50.53478</td>
</tr>
</table>
</table-wrap>
<sec>
<title>3.1. R libraries used</title>
<p>The app uses a range of R libraries for its full functionality. These can be organized in terms of their use within the app in six groups. The first group is the libraries used for its creation as <italic>a java-based application</italic> and for the main layout. Two libraries are used for this purpose: shiny (<xref ref-type="bibr" rid="B11">Chang et al., 2019</xref>) and shinydashboard (<xref ref-type="bibr" rid="B12">Chang &amp; Ribeiro, 2018</xref>). The second group is used for the <italic>data wrangling</italic> and this includes pryr (<xref ref-type="bibr" rid="B68">Wickham, 2018</xref>) and tidyr (<xref ref-type="bibr" rid="B69">Wickham &amp; Henry, 2019</xref>). The third one is used for <italic>data managing and visualization of tables</italic>: DT (<xref ref-type="bibr" rid="B71">Xie, Cheng, &amp; Tan, 2019</xref>), data.table (<xref ref-type="bibr" rid="B18">Dowle &amp; Srinivasan, 2019</xref>) and rhandsontable (<xref ref-type="bibr" rid="B47">Owen, 2018</xref>). The fourth group of libraries is used for <italic>spatial calculations of lines and contour intersections</italic>: raster (<xref ref-type="bibr" rid="B25">Hijmans, 2019</xref>), sp (<xref ref-type="bibr" rid="B48">Pebesma &amp; Bivand, 2005</xref>), gss (<xref ref-type="bibr" rid="B22">Gu, 2020</xref>), and rgdal (<xref ref-type="bibr" rid="B9">Bivand, Keitt, &amp; Rowlingson, 2019</xref>). The fifth group of libraries is used for the main <italic>visualization</italic> functionality, which includes ggplot2 (<xref ref-type="bibr" rid="B67">Wickham, 2016</xref>), ggrepel (<xref ref-type="bibr" rid="B58">Slowikowski, 2019</xref>), plotly (<xref ref-type="bibr" rid="B57">Sievert, 2018</xref>), highcharter (<xref ref-type="bibr" rid="B31">Kunst, 2019</xref>), and rAmCharts (<xref ref-type="bibr" rid="B64">Thieurmel, Marcelionis, Petit, Salette, &amp; Robert, 2019</xref>). The last group of libraries is used for the <italic>widgets layouts and display colors</italic>: shinyjs (<xref ref-type="bibr" rid="B5">Attali, 2018</xref>), shinyjqui (<xref ref-type="bibr" rid="B63">Tang, 2019</xref>), colourpicker (<xref ref-type="bibr" rid="B4">Attali, 2017</xref>), shinyBS (<xref ref-type="bibr" rid="B6">Bailey, 2015</xref>), RColorBrewer (<xref ref-type="bibr" rid="B45">Neuwirth, 2014</xref>), and wesanderson (<xref ref-type="bibr" rid="B53">Ram &amp; Wickham, 2018</xref>).</p>
</sec>
<sec>
<title>3.2. App structure</title>
<p>This section presents the general structure of the app. The app is a single web page organized by a navigation bar with all tabs available in one window. This enables going back and forward and alternating between different stages in the analysis. The motivation is to make the analysis more holistic. Traditionally, wrangling, visualization, and analysis are done at separate and distinctive stages. But with this approach, the layout allows for switching between the different stages as needed. The app has 12 main sections and they can be divided into six subsections: data and overview, visualization, graphics, gridlines, analysis, and SSANOVA. For ease of visualization, screenshots of these sections are found in the Appendix.</p>
<sec>
<title>3.2.1. Data and overview</title>
<p>There are five tabs in this section (See Figure <xref ref-type="fig" rid="F3">3</xref>). First is the <italic>Home</italic> tab, which shows a static logo and general information. The second is the <italic>Documentation</italic> tab and it offers general information on the different parts of the app. The third <italic>Load</italic> tab is where the user loads the data for analysis. The app only reads <italic>csv</italic> and <italic>tab-separated</italic> files. The data has to be previously wrangled from the specific source software so it can be read in. The fourth one is the <italic>Data</italic> tab. It shows the data imported in a table giving information on the number of speakers, segments, repetitions, and frames. The fifth tab is <italic>Manage Data</italic>. Different from the <italic>Data</italic> tab, in the <italic>Manage Data</italic> tab users can edit the data within the app. Here, the user has three options: delete, modify, and compare. In the <italic>delete</italic> option, users can delete one or more frames, repetitions, and speakers. The second option is the modify option, which can be used to relabel speaker names, segments, or repetition labels. The compare option makes comparisons of tongue contours of the selected rows. This is relevant when some data is being considered for deletion; having the option to visualize it beforehand is very important.</p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Structure of the main visualization tab.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g3.png"/>
</fig>
</sec>
<sec>
<title>3.2.2. Visualization</title>
<p>The visualization section has three parts: <italic>Data inspection, Landmarks</italic>, and <italic>In context</italic>. The first one is the <italic>Data inspection</italic> tab. Here users can explore all the data imported by allowing extensive interactions with all contours using the widgets available. This section as well as all tabs that visualize data have the structure as shown in Figure <xref ref-type="fig" rid="F3">3</xref>. The main controls on the plotted data are located on the left pane. This includes speakers, repetitions, and frames, as well as an option to save the current plot. The main plot is located on the top-right section. All the changes applied on the left controls are automatically updated on this main plot. Here, the visualization is done in individial speakers, that is, no speakers can be compared in the same graphics. One important feature in this section is that the app enables users to manually set anchor points in the visualization. Anchor points are fixed points in the figure which become the origin points from which polar-like lines are drawn. This allows for exploring areas that show more articulatory behaviour, which is the basis of the visualization and analysis approach of the app. After setting the anchor point, a second point is created by the user to create a line. The algorithm then automatically calculates distances based on the intersections between tongue contours and the line created. The app gives <italic>xy</italic> coordinates for each intersection. It also gives the angles to the left and to the right, which can be used to inspect articulatory advancement or retraction of specific segments based on angle differences (<xref ref-type="bibr" rid="B50">Proctor, 2009</xref>). The overlaid features are located on the bottom right. These are extra options to inspect the location of the point of origin and the option to manipulate the behaviour of other features such as lines, points, and labels.</p>
<p>The next visualization tab is <italic>Landmarks</italic>. This section allows for creating labels for articulatory landmarks to be analyzed in the data. These landmarks are not automatically created but instead they are manually established by the users. It can range from one single landmark to multiple landmarks. Users also have the option to modify the labels of the landmarks once these have been created. Landmark definition is not required for the visualization functionality of the app. However, they are required, as well as their assignation, for the dynamic analysis functionality.</p>
<p>The next tab is the <italic>In Context</italic> tab. This section assigns contours to the pre-established articulatory landmarks. This is a tagging capability of the app. Similar to the <italic>Data inspection</italic> tab, users can create lines with anchor points. One difference is that only one token with its corresponding repetition can be visualized, this is, no multiple repetitions are allowed in the same plot. The purpose is to establish the context in which each contour is located, allowing the user to see the previous and following contour(s). In the case of consonants, this allows for specification of the frame that contains the <italic>maximum constriction</italic>, which is defined as the frame where the constriction is held and then returns to a lower/resting position. Users can define the number of frames in context to be visualized.</p>
</sec>
<sec>
<title>3.2.3. Graphics</title>
<p>The <italic>Graphics</italic> tab gives extensive manipulation options to users. Users can edit the graphics of tongue contours, the palate trace, and the text on the images. For tongue contours and the palate trace, options include line colour, line thickness, alpha value, the smoothness level, and the line type. The last option is to modify the text in the plots. It includes font type, font size, and colour of the axis labels, axis ticks, and legend. The settings modified in this section automatically apply to plots in the tabs <italic>Data Inspection, Landmarks</italic>, and <italic>In Context</italic>.</p>
<p>In relation to the smoothness parameter, this is only used within the app for visualization purposes. The smoothness is therefore used for the front end to smooth tongue contours in the plots. This does not affect the calculations of the intersections as explained in Section 3.3. For the visualization, users can change the smoothness method of contours in the <italic>Graphics</italic> section. The methods implemented are gam and loess, as they are available in the ggplot2 (<xref ref-type="bibr" rid="B67">Wickham, 2016</xref>) R package.</p>
</sec>
<sec>
<title>3.2.4. Gridlines</title>
<p>The <italic>Gridlines</italic> tab is a foundational section in the app (this section is expanded in Section 3.3). This section is used to establish all gridlines in the data. This can be done on one or multiple segments at a time. The first step is to define the location of the origin point. The first option is the manual option where the user defines where the best location is by clicking on the image. The other two options are for automatic location of the origin point, which is described in more detail in Section 3.3.1. The first option is the Wide setting. It defines the origin point considering the most anterior initial point and the most posterior last point of contours. On the other hand, the Narrow option considers the most posterior initial point and the most anterior last point. The following step after establishing the origin point is the definition of the analysis fan view, which establishes the range of comparison across contours. There are four options. The first three are the same as in the origin point, Manual, Narrow, and Wide. The extra option is the Angle option. It allows the user to define the analysis fan view based on angle degrees for the angle aperture. Once the field-view is established, the user defines the number of gridlines of the fan. The two options are by number and by angle. If chosen by numbers, the user selects the number of gridlines. If by angle, the user decides to choose the location of the grid lines by angle increments.</p>
</sec>
<sec>
<title>3.2.5. Analysis</title>
<p>The last tab is <italic>Analysis</italic>. In this section, the user has the opportunity to examine the data established in the previous sections. It is based on the intersections calculated in the <italic>Gridlines</italic> tab and with the landmarks created in the <italic>Landmarks</italic> tab and tagged in the <italic>In Context</italic> tab. Following the same layout structure as in the previous sections, the main controls are on the left and the main plot is at the top-right. The main difference is that there is an extra visualization at the bottom right that shows the dynamic differences between the contours selected. The four options are displacement, distance, velocity, and acceleration, which are explained in Section 3.3.4.</p>
</sec>
</sec>
<sec>
<title>3.3. Analysis baseline</title>
<p>The analysis baseline is centred on a gridlines approach developed in Gonzalez (<xref ref-type="bibr" rid="B21">2015</xref>), which is similar to the one used in AAA (<xref ref-type="bibr" rid="B70">Wrench, 2012</xref>) and other studies (cf. <xref ref-type="bibr" rid="B35">Liker, Zori&#263;, Zharkova, &amp; Gibbon, 2019</xref>; <xref ref-type="bibr" rid="B62">Strycharczuk &amp; Scobbie, 2015</xref>). These gridlines are a composition of multiple lines with the same origin point and projected in different angles to create a grid-like analysis fan view. The location of the gridlines can be determined purely on data-internal events, this is, not based on anatomical parameters but on contour behaviour. The gridlines origin is fixed for each participant. The use of gridlines is implemented to capture articulatory activity at different locations in tongue ultrasound images as well as to measure gestures in specified articulatory locations. The measurement values are defined from the intersection points between gridlines and tongue contours. Since the gridlines are fixed for all realizations of a given speaker, all the differences can therefore be interpreted as pertaining to differences in the articulatory patterns shown in the tongue contours.</p>
<p>This type of approach allows analysis of data in two dimensions. The first is the kinaesthetic dimension, which examines the data from static (e.g., comparing the maximum constriction points in two different segments, as observed in Figure <xref ref-type="fig" rid="F4">4</xref>) and dynamic perspectives (e.g., comparing how two contours change across time). The second dimension is the location of the articulatory activity, this is, analyzing contours either at a general level (e.g., comparing the articulatory activity between two segments along the full length of their surfaces), or at a specific level (e.g., comparing articulatory activity between two segments only in the tongue body section). This gridline approach can be used to carry out three types of analysis in ultrasound studies: temporal, spatial, and spatial-temporal. Temporal analyses measure time differences in the articulation between contours, for instance, which segment takes less time to reach its maximum constriction point. For the spatial analysis, the program can carry out spatial differences between contours, for example, the tongue movement from a lower position to a higher position in the vocal tract as done by displacement analyses. And the last one, spatial-temporal, carries out velocity and acceleration measurements to compare segments of interest in either whole full trajectories or isolated areas of the tongue across time. The following sections describe the three most important components of the gridlines analysis: the origin point, the analysis fan view, and the intersections.</p>
<fig id="F4">
<label>Figure 4</label>
<caption>
<p>Layout of the Analysis tab.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g4.png"/>
</fig>
<sec>
<title>3.3.1. Gridlines origin location</title>
<p>The definition of the gridlines origin location can be done manually or automatically. The first option allows the user to locate the origin point manually by clicking on the desired location within the plot. The second is the automatic option. In this case there are two further options, selection of either a narrow window or a wide window. In the Wide option, the <italic>x</italic> value of the midpoint is located between the extreme points: the most advanced point and the most retracted point of tongue contours, as shown in Figure <xref ref-type="fig" rid="F5">5</xref>. For the Narrow option, the <italic>x</italic> value is calculated between the most retracted first point and the most advanced last point of tongue contours.</p>
<fig id="F5">
<label>Figure 5</label>
<caption>
<p>Definition of the x value of the Origin Point.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g5.png"/>
</fig>
<p>The <italic>y</italic> value is calculated by first extracting the highest <italic>y</italic> value of all contours and the lowest <italic>y</italic> value of all contours, as shown as the highest and lowest points in Figure <xref ref-type="fig" rid="F6">6</xref>. Then the final <italic>y</italic> value is calculated by subtracting the highest <italic>y</italic> value from the lowest <italic>y</italic> value.</p>
<fig id="F6">
<label>Figure 6</label>
<caption>
<p>Definition of the y value of the Origin Point.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g6.png"/>
</fig>
</sec>
<sec>
<title>3.3.2. Analysis Fan View</title>
<p>After the gridlines origin has been created, the next step is to define the Analysis Fan View (AFV), which establishes how much of each tongue contour is included in the analysis for comparison. A wider AFV can capture more articulatory activity at different sections of the tongue. However, if it is too wide, it risks having sections where not all contours are comparable due to tokens which are missing intersections in specific areas. On the other hand, a narrow AFV can be used to isolate areas of interest and analysis. The risk is that if looking at a very narrow AFV, the analysis may lose important articulatory activity that is outside the range. The trade-off between the two therefore has to be kept in mind by the researcher. The other sections of the app, especially <italic>Data Inspection</italic> and <italic>In Context</italic>, allow users to inspect areas of interest before deciding the AFV for analysis.</p>
<p>Similar to the gridlines origin, the AFV can be established manually or automatically. For the manual option, the user can choose the anterior and posterior lines manually by clicking on the plot or specifying the aperture angle for each of them. In the case of automatic options, it can be narrow or wide. These two depend on the angle aperture (See Figure <xref ref-type="fig" rid="F7">7</xref> for reference). In the case of the wide option, the left line is located at the intersection point with the widest angle. The right line is located at the intersection point with the narrowest angle. The main strength of this option is that is captures all contours across their length.</p>
<fig id="F7">
<label>Figure 7</label>
<caption>
<p>Creating the Analysis Fan View.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g7.png"/>
</fig>
<p>The narrow option creates an AFV that captures only the sections that are within the common area. There are two stages to calculate the second point for each line. For the left line, the algorithm calculates the angle between the origin point and the last point of each contour. It first starts with the widest angle and for each iteration it checks whether the projected line intersects with all contours. If the projected line does not intersect with all contours, the following angle is selected. The process stops when the line intersects with all contours in the data. The same process is applied to find the right line, but in this case, it starts from the narrowest angle and iterates until finding a line that intersects with all tongue contours. The result is then an angle aperture which captures a section that is common in all the data. As shown in the triangle in Figure <xref ref-type="fig" rid="F7">7</xref>, there is a common area which captures all contours that stay within the AFV. The areas at the left and the right outside the common area are not considered in the analysis.</p>
</sec>
<sec>
<title>3.3.3. Gridlines definition</title>
<p>When the AFV is established, the next step is the number of gridlines, which is decided based on either a selected number of gridlines or angle increments. In the first option, the user defines a specific number of gridlines and the algorithm divides the angle aperture of the AFV by the number of lines specified. For the angle option, the user defines the angle step increment. Then the lines are added by the angle step increment, starting from the right line and adding new lines until reaching the maximum angle within the field view aperture. In both cases, the result is a fan-like view of gridlines superimposed on all tongue contours. This allows extraction of tongue contours based on intersections between traces and gridlines. Figure <xref ref-type="fig" rid="F8">8</xref> shows the resulting gridlines with their corresponding angles in a sample data.</p>
<fig id="F8">
<label>Figure 8</label>
<caption>
<p>Final Gridlines in the data, Wide option at the left and Narrow option at the right.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g8.png"/>
</fig>
</sec>
<sec>
<title>3.3.4. Intersection calculations</title>
<p>The following step is the intersection points. After all gridlines have been established, all intersections are calculated for all contours. The result is a new data frame with the same number of tongue contours but narrowed down to the intersection points within the AFV. This new data is the baseline for the analysis within the app. The app can analyze dynamic patterns in four ways: displacement, distance, velocity, and acceleration, which are key when examining tongue motion. In this context, tongue motion is defined as the change of a tongue section within the mouth in respect to time (how fast the tongue is moving). One parameter is displacement, which is the length of the path traveled by the tongue section from one landmark to another, based on a specific gridline. It is important to point out that this displacement is a relative measure from one tongue contour to another. Different from EMA techniques, which track tongue flesh points, the displacement calculated here can only measure the relative movement from point A to point B in a given gridline. It is represented in Figure <xref ref-type="fig" rid="F9">9</xref> and calculated using Formula 1 where d<sub>LMB</sub> is the distance from the origin point to the intersection with a tongue contour in the second articulatory landmark and d<sub>LMA</sub> is the distance from the origin point to the intersection with a tongue contour in the first articulatory landmark.</p>
<fig id="F9">
<label>Figure 9</label>
<caption>
<p>Displacement calculation between landmarks.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g9.png"/>
</fig>
<p><bold>Formula 1:</bold> Displacement formula.</p>
<disp-formula id="FD1">
<alternatives>
<mml:math id="Eq001-mml"><mml:mrow><mml:mi>d</mml:mi><mml:mo>=</mml:mo><mml:msub><mml:mi>d</mml:mi><mml:mrow><mml:mi>LMB</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>d</mml:mi><mml:mrow><mml:mi>LMA</mml:mi></mml:mrow></mml:msub></mml:mrow></mml:math>
<tex-math id="M1">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
d = {d_{LMB}} - {d_{LMA}}
\]
\end{document}
</tex-math>
<graphic xlink:href="labphon-12-6463-e1.gif"/>
</alternatives>
</disp-formula>
<p>It first calculates the distances from the origin point to the intersections of the first landmark contours at a given gridline. For example, the displacement of the tongue from the midpoint of a vowel to the maximum constriction of a following consonant on the 5<sup>th</sup> gridline. The displacement can therefore be positive, negative, or zero. Displacement in this framework is defined as the movement of the tongue in specific locations, which are determined by the gridlines. For each line, the displacement in mm or pixels is the distance from the first tongue contour of comparison (Intersection A in Figure <xref ref-type="fig" rid="F9">9</xref>) to the next articulatory moment (Intersection B in Figure <xref ref-type="fig" rid="F9">9</xref>) to capture more fined-tuned articulatory patterns.</p>
<p>The second dynamic analysis is distance. It calculates all the space the tongue has covered throughout its trajectory from the first landmark to the second landmark. Figure <xref ref-type="fig" rid="F10">10</xref> depicts this difference. On the left, there are three tongue intersections and they ascend in the gridline from the first intersection to the third, both landmarks. On the right, there are four tongue intersections. One difference is that from the first to the second, there is a downward movement, then it ascends to the third and finally to the fourth. As shown here, if we consider only landmark displacement, they have the same value. However, the distances for each token differ, since the example at the right has a transition contour intersection that adds more distance to the trajectory. This difference between displacement and distance is important when considering multiple contours in the trajectory between articulatory landmarks. This is also implemented in the app.</p>
<fig id="F10">
<label>Figure 10</label>
<caption>
<p>Difference between Displacement and Distance.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g10.png"/>
</fig>
<p>Velocity, the third dynamic analysis, measures the speed of a section of the tongue in a specified gridline (how fast the displacement is in relation to time). In Strycharczuk and Scobbie (<xref ref-type="bibr" rid="B62">2015</xref>), tongue contour velocity is measured based on upward or downward movements along lines placed on a fan-like shape. This measurement is also implemented here and is calculated following Formula 2, where <italic>t</italic> represents time calculated by multiplying the number of frames between the first and second landmark by the frame rate of the ultrasound images (e.g., 30 fps are 0.033 seconds per frame-to-frame transition).</p>
<p><bold>Formula 2:</bold> Velocity formula.</p>
<disp-formula id="FD2">
<alternatives>
<mml:math id="Eq002-mml"><mml:mrow><mml:mi>v</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msub><mml:mi>d</mml:mi><mml:mrow><mml:mi>LMB</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>d</mml:mi><mml:mrow><mml:mi>LMA</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mi>t</mml:mi></mml:mfrac></mml:mrow></mml:math>
<tex-math id="M2">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
v = \frac{{{d_{LMB}} - {d_{LMA}}}}{t}
\]
\end{document}
</tex-math>
<graphic xlink:href="labphon-12-6463-e2.gif"/>
</alternatives>
</disp-formula>
<p>The last metric, acceleration, measures changes in the velocity of that section, with respect to time, using Formula 3.</p>
<p><bold>Formula 3:</bold> Acceleration formula.</p>
<disp-formula id="FD3">
<alternatives>
<mml:math id="Eq003-mml"><mml:mrow><mml:mi>a</mml:mi><mml:mo>=</mml:mo><mml:mfrac><mml:mrow><mml:msub><mml:mi>d</mml:mi><mml:mrow><mml:mi>LMB</mml:mi></mml:mrow></mml:msub><mml:mo>&#x2212;</mml:mo><mml:msub><mml:mi>d</mml:mi><mml:mrow><mml:mi>LMA</mml:mi></mml:mrow></mml:msub></mml:mrow><mml:mrow><mml:msup><mml:mi>t</mml:mi><mml:mn>2</mml:mn></mml:msup></mml:mrow></mml:mfrac></mml:mrow></mml:math>
<tex-math id="M3">
\documentclass[10pt]{article}
\usepackage{wasysym}
\usepackage[substack]{amsmath}
\usepackage{amsfonts}
\usepackage{amssymb}
\usepackage{amsbsy}
\usepackage[mathscr]{eucal}
\usepackage{mathrsfs}
\usepackage{pmc}
\usepackage[Euler]{upgreek}
\pagestyle{empty}
\oddsidemargin -1.0in
\begin{document}
\[
a = \frac{{{d_{LMB}} - {d_{LMA}}}}{{{t^2}}}
\]
\end{document}
</tex-math>
<graphic xlink:href="labphon-12-6463-e3.gif"/>
</alternatives>
</disp-formula>
<p>Both velocity and acceleration are based on tongue displacement: Velocity measures the rate of change of tongue displacement and acceleration measures the rate of change of the velocity. One important aspect of interpreting positive and negative values is that it depends on the measurement analyzed. In the case of displacement, distance, and velocity, negative values represent negative motion in relation to the baseline of comparison, which in this case is the first articulatory landmark. This reflects the directionality in the specified gridline. In the case of acceleration, it depends on the velocity. Since acceleration measures changes in velocity, negative acceleration takes place when there is slowing motion from one frame to another. In these terms, the first initial acceleration, from the first landmark contour to the next frame, the acceleration is positive. Then the following values can be positive or negative, depending on whether velocity slows down in frame sequences. This allows measuring a very important aspect of tongue articulations not observed at the spatial level but on a spatial-temporal dimension.</p>
</sec>
</sec>
<sec>
<title>3.4. Comparing multiple segments with dynamic metrics</title>
<p>The tongue displacement/distance analyses examine spatial trajectories along the whole analysis fan view and the velocity/acceleration analyses look at this displacement across time. The purpose is to capture patterns of tongue gestures along extended or isolated sections of the tongue. Multiple time windows are required for this analysis, with at least two landmarks for comparison. In the example below, I present displacement between three articulatory landmarks, namely, Previous Vowel (PV), Maximum Constriction (MC) of an obstruent consonant, and Following Vowel (FV), in the case of a VCV sequence. There are therefore two transitions, one from PV to MC (PV-MC), and another from MC to FV (MC-FV).</p>
<p>The analysis takes the distance intersections of the first moment as the baseline for comparison. In the case of the PV-MC displacements, the distance between the origin point and the PV intersections are the baseline. The baseline for the MC-FV displacement is the distance intersections for MC. These baseline distance intersections are compared to the distance intersections of the second moment. Figure <xref ref-type="fig" rid="F11">11a</xref> and Figure <xref ref-type="fig" rid="F11">11b</xref> show both moments. Point 1.1 in <italic>a</italic> shows the baseline intersection, which is the intersection in PV. The first distance calculated is the distance from the origin to the 1.1 intersection. The second distance calculated is from the origin to the 1.2 intersection. The origin-PV distance is compared to the origin-MC distance. Then the difference is calculated, which is the displacement distance in the given gridline.</p>
<fig id="F11">
<label>Figure 11</label>
<caption>
<p>Positive and Negative Displacements representations.<xref ref-type="fn" rid="n3">3</xref></p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g11.png"/>
</fig>
<p>In addition to the displacement distance, the analysis also examines the orientation of the displacement. If the distance in a gridline is longer for PV and shorter for MC, then it is classified as a negative displacement (as seen on the difference between point 40.1 and 40.2 on <italic>a</italic>). On the contrary, if the distance from the origin to the MC intersection is longer than the PV intersection, then it is classified as positive displacement (as seen on the difference between point 1.1 and 1.2 on <italic>a</italic>). This type of analysis allows measuring the directionality of sections of the tongue for achieving its articulatory target. After the displacement is calculated for each contour, the result is a displacement graph across all gridlines, as represented in Figure <xref ref-type="fig" rid="F12">12</xref>.</p>
<fig id="F12">
<label>Figure 12</label>
<caption>
<p>Representation of a displacement graph.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g12.png"/>
</fig>
<p>The comparison between segments can then be achieved with the displacement graph approach. Figure <xref ref-type="fig" rid="F13">13</xref> represents the process of comparison. In <italic>a</italic> and <italic>b</italic>, displacements are calculated for each segment. As an example, Segment A displacement shows strong positive activity at the body and back sections of the tongue, whereas there is negative activity in the front section, but less than the positive displacement. On the other hand, Segment B does not display as much activity at the back of the tongue compared to Segment A and with little positive displacement at the first gridlines. When these are compared on the superimposed figure in <italic>c</italic>, the displacement shows two main distinctive differences. At the front section of the tongue, Segment B has positive displacement whereas Segment A shows negative movement. At the back section, even though both show positive displacement, Segment A shows more movement than Segment B. These differences can be assessed qualitatively as well as quantitatively.</p>
<fig id="F13">
<label>Figure 13</label>
<caption>
<p>Representation of displacement comparisons between two segments.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g13.png"/>
</fig>
</sec>
<sec>
<title>3.5. Smoothing splines analysis of variance</title>
<p>The app offers a functionality to analyze tongue contours using the Smoothing Splines Analysis of Variance (henceforth SSANOVA) approach. This section is implemented following the procedure developed in Davidson (<xref ref-type="bibr" rid="B14">2006</xref>) and Mielke (<xref ref-type="bibr" rid="B40">2015</xref>). In this section, the first step is to fit the data using smoothing splines by fitting a polynomial function connecting the tongue contour points. This type of approach allows the user to measure multiple whole contour trajectories and analyze how repetitions of different sound segments compare to each other. When using SSANOVAs, the common practice when comparing tongue contours is to examine the Bayesian confidence intervals that are constructed from the multiple repetitions. If the confidence intervals of the two compared groups overlap, then this section is described as not having significant differences. On the other hand, when there is no overlap in the confidence intervals, then this section shows significant differences between the groups.</p>
<p>The advantage of this approach is that it allows users to examine tongue contours not as a whole, but to focus only on the specific areas that display differences between groups. This type of approach is in line with the nature of the tongue, which can display articulatory similarities between segments in one section but not in another. Therefore, by implementing SSANOVAs, the app offers another analysis tool to examine significant differences both visually and statistically sound.</p>
<p>In this section of the app, see Figure <xref ref-type="fig" rid="F14">14</xref>, the user selects the speaker and the two groups to compare. In Figure <xref ref-type="fig" rid="F14">14</xref>, the user compares all the repetitions for /s/ with all the repetitions for /&#643;/. There are three outputs, namely, the individual contours, the overall comparisons, and the confidence intervals. All of these, along with the numeric data, can be downloaded using the corresponding graphical user interfaces.</p>
<fig id="F14">
<label>Figure 14</label>
<caption>
<p>SSANOVA tab options.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g14.png"/>
</fig>
<p>For the individual contours, as observed in Figure <xref ref-type="fig" rid="F15">15</xref>, all raw tongue contours are plotted. This helps to inspect the input data which is then used in the analysis.</p>
<fig id="F15">
<label>Figure 15</label>
<caption>
<p>First visual output in the SSANOVA analysis. This shows the individual contours in each group.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g15.png"/>
</fig>
<p>For the overall comparison, the splines from the analysis are displayed (see Figure <xref ref-type="fig" rid="F16">16</xref>). These show the best fit for each group as well as the confidence intervals across the length of the trajectories.</p>
<fig id="F16">
<label>Figure 16</label>
<caption>
<p>Second visual output displaying the smoothed splines for each group.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g16.png"/>
</fig>
<p>The last tab shows the Bayesian confidence intervals for each group. The horizontal dashed line at 0 represents the baseline to identify whether areas along the trajectories are significantly different (see Figure <xref ref-type="fig" rid="F17">17</xref>). The areas where trajectories touch the 0 line represent sections that are not significantly different, for example, the cross-sections right at the beginning and end of both panels and around 20% in from left to right. The other non-overlapping sections are significantly different between the groups. This shows that the area around 60% into the trajectories is the area with most significant differences. This section corresponds to the tongue body differences between /s/ and /&#643;/. The fact that /&#643;/ is lower than /s/ is an artefact of the pixel measurement, which is inverted: Lower values correspond to higher positions of the tongue.</p>
<fig id="F17">
<label>Figure 17</label>
<caption>
<p>Third visual output showing the Bayesian Confidence Intervals for each group.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g17.png"/>
</fig>
<p>These three visual aids, along with the numeric output, give users a powerful analysis tool that can be used for effective tongue contour analysis. In this way, the widely used SSANOVA approach can also be used in the app.</p>
</sec>
</sec>
<sec>
<title>4. Demonstration study</title>
<p>In this section, I present a sample test from a subset of the data obtained in Gonzalez (<xref ref-type="bibr" rid="B21">2015</xref>). For demonstration purposes, I only illustrate two segments (/s/ and /&#643;/ in English) from one speaker. Each segment has five repetitions and they appear in the carrier sentence, <italic>Please utter X publicly</italic>, where X is the carrier word, <italic>sack</italic> (for /s/) and <italic>shack</italic> (for /&#643;/). Each token appears between two vowels, /&#601;/ and /&#230;/. Figure <xref ref-type="fig" rid="F18">18</xref> shows all the tokens and the gridlines established with the Narrow option to only focus on the common area of all contours. Two articulatory landmarks were established, Previous Vowel and Maximum Constriction with 20 gridlines. All results shown below are based on the intersections from the gridlines.</p>
<fig id="F18">
<label>Figure 18</label>
<caption>
<p>Sample data tokens and gridlines.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g18.png"/>
</fig>
<p>One hypothesis tested is that for PV-MC transitions the palato-alveolar segment /&#643;/ shows more positive displacement of the middle section of the tongue when compared to alveolar /s/, which is hypothesized to show positive displacement at the most advanced sections of the tongue to achieve the MC at the alveolar region. The analysis does not only focus on positive displacement patterns but also on negative displacements. Figure <xref ref-type="fig" rid="F19">19</xref> shows the raw plots of both segments at the MC point, /s/ shown with the solid line and /&#643;/ shown with the dashed line. The first observation is that there is more raising of /&#643;/ at the tongue body. This shows that the main difference between /s/ and /&#643;/ is that /&#643;/ has more articulatory activity at the body apex. There are no strongly distinctive differences at the front section of the tongue.</p>
<fig id="F19">
<label>Figure 19</label>
<caption>
<p>Raw tongue contours from the gridlines intersections.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g19.png"/>
</fig>
<p>I added another type of visualization which uses heatmaps to observe articulatory activity. This is available in the <italic>Analysis</italic> tab. The MC heatmaps of /s/ and /&#643;/ are shown in Figure <xref ref-type="fig" rid="F20">20</xref>. The darker the area, the more activity there is at that specific location. As observed in the comparison, both have strong activity at the tongue body apex. The difference is that /&#643;/ shows more localized activity at that section, whereas /s/ also has more activity spreading mainly at the front section of the tongue. The heatmap visualization then adds another perspective for analyzing articulatory activity.</p>
<fig id="F20">
<label>Figure 20</label>
<caption>
<p>Heatmaps during MC across all five repetitions.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g20.png"/>
</fig>
<p>The next step is to inspect each segment individually to observe displacement differences. First, I inspect /&#643;/ in Figure <xref ref-type="fig" rid="F21">21</xref>. In the left I see all the five tokens at two landmarks; the solid lines are the MC contours and the dashed lines are the PV contours. The hue of the contours represents the repetition number: Darker hues are earlier repetitions. On the right, all repetitions per landmark have been grouped, which shows the same pattern as the individual tokens: MC tokens are raised and more advanced at the tongue body apex than the PV contours.</p>
<fig id="F21">
<label>Figure 21</label>
<caption>
<p>Contour differences from PV to MC for /&#643;/. Shaded areas in the right plot represent confidence intervals.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g21.png"/>
</fig>
<p>The patterns for /s/ are different (See Figure <xref ref-type="fig" rid="F22">22</xref>), with the individual tokens at the left and the grouped ones at the right. The patterns observed in /s/ differ in that there is lowering at the body apex from PV to MC. There is also lowering at the front section of the tongue from PV to MC. Finally, another difference is that the lowest section of the tongue at the tongue back shows less movement for /s/ than for /&#643;/.</p>
<fig id="F22">
<label>Figure 22</label>
<caption>
<p>Contour differences from PV to MC for /s/.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g22.png"/>
</fig>
<p>Figure <xref ref-type="fig" rid="F23">23</xref> presents the heatmap transitions from PV to MC, which show more local differences. As observed in the contour plots, activity at the body apex is distinctive, with /&#643;/ having more raised and advanced articulations. In the case of /s/, there is more significant gestural movement at the front section of the tongue than for /&#643;/. This reflects the expected behaviour for /s/: Since its MC is achieved at the alveolar ridge, it shows more activity at the very front sections of the contours. In the case of /&#643;/, more activity is expected at the body apex, which is the section of the tongue to achieve the MC by approaching the palato-alveolar region.</p>
<fig id="F23">
<label>Figure 23</label>
<caption>
<p>Heatmaps from PV to MC across all repetitions.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g23.png"/>
</fig>
<p>The last type of analyses can be used to observe dynamic patterns in a deeper way. First, we analyze the velocity patterns of /s/ in red and /&#643;/ in highlighted blue in Figure <xref ref-type="fig" rid="F24">24</xref>. Velocity values are shown across all 20 gridlines and grouped per segment.</p>
<fig id="F24">
<label>Figure 24</label>
<caption>
<p>Velocity patterns from PV to MC.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g24.png"/>
</fig>
<p>The comparison shows four main areas of activity. At the front section, from gridlines 1 to 3, there are positive velocities for /s/, whereas there are negative velocities for /&#643;/ between gridlines 4 and 7. The positive velocities for /&#643;/ that are comparable to /s/ are from gridlines 9 and 13. Gridlines 15 to 20 show negative velocities for both. Since velocities are closely related to the distances, we can interpret contour movement related to the time it takes to achieve the articulatory targets. First, regardless of their positive or negative values, larger velocities correspond to those areas where the movement of the tongue moves faster from PV to the MC. One key finding here is that both segments show the most prominent positive velocities exactly in those areas where the main constriction is targeted in the vocal cavity. In the case of /s/, it is at the first gridlines that would correspond to the alveolar region, where the MC is expected to take place coming from the previous vowel. In the case of /&#643;/, the gridlines with more prominent positive velocities correspond to the postalveolar region, where its MC takes place.</p>
<p>Acceleration results also show important patterns in Figure <xref ref-type="fig" rid="F25">25</xref>. As noted above, negative acceleration values correspond to transitions where there is slowing down in the articulation. Again, acceleration values are grouped across segments throughout the 20 gridlines with /s/ in red and /&#643;/ in blue.</p>
<fig id="F25">
<label>Figure 25</label>
<caption>
<p>Acceleration patterns from PV to MC.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-12-6463-g25.png"/>
</fig>
<p>Similar to velocity results, there is more prominent acceleration at the front section for /s/ and in the middle section for /&#643;/. It shows that there is strong extended acceleration also at the tongue body, where /&#643;/ achieves its MC. The pattern for /&#643;/ does not show strong acceleration at the front section of the tongue, except for gridline 2. Also, unlike velocity measurements, acceleration here shows that there is not much activity happening at the back section of the tongue, especially for /&#643;/. Finally, this analysis is in line with the heatmap plots in Figure <xref ref-type="fig" rid="F23">23</xref>. Both show that the articulatory activity of /s/ is more spread across the vocal cavity, whereas it is more localized for /&#643;/, at the middle section. Since the articulation of the palato-alveolar segment requires significant movement of the tongue body, energy shown in the acceleration suggests that most of the muscle effort is focused on moving the tongue body, leaving the back and front areas of the tongue less dynamic. On the other hand, since to achieve its MC /s/ mainly requires movement at the front section of the tongue, which has less mass than the tongue body, the muscle effort can be spread in other sections of the tongue, that is, not putting all the effort into movement of the front section, as it is done in /&#643;/ to move the tongue body.</p>
</sec>
<sec>
<title>5. Conclusions</title>
<p>Articulatory analysis of speech phenomena has helped us have a better understanding of how articulation and gestures function. The use of technologies like ultrasound has been of undeniable importance in the field. By developing technologies like the one described in this paper, we can expand our study of human articulation, not only in normal speech but also in cases of speech impediment and speech development. It is my aim then that by implementing this tool, many of the hurdles for efficient tongue ultrasound analysis can be overcome so we can tackle new challenges of research. Being an open source tool, UVA opens the door to broadening the scope of ultrasound studies by allowing users to create new analysis algorithms and then implement them as needed within the app from the source code.</p>
</sec>
<sec sec-type="supplementary-material">
<title>Additional File</title>
<p>The additional file for this article can be found as follows:</p>
<supplementary-material id="S1" xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="https://doi.org/10.16995/labphon.6463.s1">
<!--[<inline-supplementary-material xlink:title="local_file" xlink:href="labphon-12-6463-s1.pdf">labphon-12-6463-s1.pdf</inline-supplementary-material>]-->
<label>Appendix</label>
<caption>
<p>PDF file containing the User Interface for all the main components of the app. DOI: <uri>https://doi.org/10.16995/labphon.6463.s1</uri></p>
</caption>
</supplementary-material>
</sec>
</body>
<back>
<fn-group>
<fn id="n1"><p>In some studies, tongue contours are not extracted but rather the whole images are analyzed, as in Mielke et al. (<xref ref-type="bibr" rid="B41">2017</xref>).</p></fn>
<fn id="n2"><p>For a detailed description of specific aspects in ultrasound studies see Stone (<xref ref-type="bibr" rid="B59">2005</xref>).</p></fn>
<fn id="n3"><p>Figure <xref ref-type="fig" rid="F11">11</xref> presents the palate trace contour. This version of the app does not have the capability to measure distances from tongue contours to palate traces. However, this is a feature that I am planning to implement in future versions.</p></fn>
</fn-group>
<ack>
<title>Acknowledgements</title>
<p>This paper and the research behind it would not have been possible without the exceptional guidance and support of my Ph.D. thesis supervisors from which the main theoretical background of this app comes: Mark Harvey, Michael Proctor, Katherine Demuth, Susan Lin, and Alan Libert. I would also like to thank Rosey Bollington for her comments and suggestions in one of the reviews of the manuscript. The app functionality has also been used in another workshop from which very insightful comments were given by Chlo&#233; Diskin-Holdaway, Deborah Loakes, Rosey Billington, Hywel Stoakes, and Sam Kirkham. Finally, I want to thank Patrycja Strycharczuk, Stefano Coretta, the anonymous reviewers of this manuscript, and the editors, for their invaluable comments and insights in the shape and content of the final manuscript. The generosity and expertise of everyone have improved this paper in innumerable ways and saved me from many errors. Those that inevitably remain are entirely my own responsibility.</p>
</ack>
<sec>
<title>Competing Interests</title>
<p>The author has no competing interests to declare.</p>
</sec>
<ref-list>
<ref id="B1"><label>1</label><mixed-citation publication-type="journal"><string-name><surname>Ahn</surname>, <given-names>S.</given-names></string-name> (<year>2015</year>). <article-title>Utterance-initial voiced stops in American English: An ultrasound study</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>138</volume>(<issue>1777</issue>). DOI: <pub-id pub-id-type="doi">10.1121/1.4933625</pub-id></mixed-citation></ref>
<ref id="B2"><label>2</label><mixed-citation publication-type="journal"><string-name><surname>Ahn</surname>, <given-names>S.</given-names></string-name> (<year>2018</year>). <article-title>The role of tongue position in laryngeal contrasts: An ultrasound study of English and Brazilian Portuguese</article-title>. <source>Journal of Phonetics</source>, <volume>71</volume>, <fpage>451</fpage>&#8211;<lpage>467</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.wocn.2018.10.003</pub-id></mixed-citation></ref>
<ref id="B3"><label>3</label><mixed-citation publication-type="confproc"><string-name><surname>Alwabari</surname>, <given-names>S.</given-names></string-name> (<year>2019</year>). <article-title>An Ultrasound Study on Gradient Coarticulatory Pharyngealization and Its Interaction with Arabic Phonemic Contrast</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>1729</fpage>&#8211;<lpage>1733</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B4"><label>4</label><mixed-citation publication-type="webpage"><string-name><surname>Attali</surname>, <given-names>D.</given-names></string-name> (<year>2017</year>). <article-title>colourpicker: A Colour Picker Tool for Shiny and for Selecting Colours in Plots (Version 1.0) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=colourpicker</uri></mixed-citation></ref>
<ref id="B5"><label>5</label><mixed-citation publication-type="webpage"><string-name><surname>Attali</surname>, <given-names>D.</given-names></string-name> (<year>2018</year>). <article-title>shinyjs: Easily Improve the User Experience of Your Shiny Apps in Seconds (Version 1.0) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=shinyjs</uri></mixed-citation></ref>
<ref id="B6"><label>6</label><mixed-citation publication-type="webpage"><string-name><surname>Bailey</surname>, <given-names>E.</given-names></string-name> (<year>2015</year>). <article-title>shinyBS: Twitter Bootstrap Components for Shiny (Version 0.61) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=shinyBS</uri></mixed-citation></ref>
<ref id="B7"><label>7</label><mixed-citation publication-type="journal"><string-name><surname>Barberena</surname>, <given-names>L. D.</given-names></string-name>, <string-name><surname>Portalete</surname>, <given-names>C. R.</given-names></string-name>, <string-name><surname>Simoni</surname>, <given-names>S. N.</given-names></string-name>, <string-name><surname>Prates</surname>, <given-names>A. C.</given-names></string-name>, <string-name><surname>Keske-Soares</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Mancopes</surname>, <given-names>R.</given-names></string-name> (<year>2017</year>). <article-title>Electropalatography and its correlation to tongue movement ultrasonography in speech analysis</article-title>. <source>Codas</source>, <volume>29</volume>(<issue>2</issue>), <elocation-id>e20160106</elocation-id>. DOI: <pub-id pub-id-type="doi">10.1590/2317-1782/20172016106</pub-id></mixed-citation></ref>
<ref id="B8"><label>8</label><mixed-citation publication-type="journal"><string-name><surname>Beare</surname>, <given-names>R.</given-names></string-name> (<year>2018</year>). <article-title>Processing of speech ultrasound data (Version 0.0.0.9000) [R Package]</article-title>.</mixed-citation></ref>
<ref id="B9"><label>9</label><mixed-citation publication-type="webpage"><string-name><surname>Bivand</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Keitt</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Rowlingson</surname>, <given-names>B.</given-names></string-name> (<year>2019</year>). <article-title>rgdal: Bindings for the &#8216;Geospatial&#8217; Data Abstraction Library (Version 1.4-4) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=rgdal</uri></mixed-citation></ref>
<ref id="B10"><label>10</label><mixed-citation publication-type="journal"><string-name><surname>Bressmann</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Koch</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Ratner</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Seigel</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Binkofski</surname>, <given-names>F.</given-names></string-name> (<year>2015</year>). <article-title>An ultrasound investigation of tongue shape in stroke patients with lingual hemiparalysis</article-title>. <source>Journal of Stroke and Cerebrovascular Diseases</source>, <volume>24</volume>(<issue>4</issue>), <fpage>834</fpage>&#8211;<lpage>839</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.jstrokecerebrovasdis.2014.11.027</pub-id></mixed-citation></ref>
<ref id="B11"><label>11</label><mixed-citation publication-type="webpage"><string-name><surname>Chang</surname>, <given-names>W.</given-names></string-name>, <string-name><surname>Cheng</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Allaire</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Xie</surname>, <given-names>Y.</given-names></string-name>, &amp; <string-name><surname>McPherson</surname>, <given-names>J.</given-names></string-name> (<year>2019</year>). <article-title>shiny: Web Application Framework for R (Version 1.3.2) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=shiny</uri></mixed-citation></ref>
<ref id="B12"><label>12</label><mixed-citation publication-type="webpage"><string-name><surname>Chang</surname>, <given-names>W.</given-names></string-name>, &amp; <string-name><surname>Ribeiro</surname>, <given-names>B. B.</given-names></string-name> (<year>2018</year>). <article-title>shinydashboard: Create Dashboards with &#8216;Shiny&#8217; (Version 0.7.1) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=shinydashboard</uri></mixed-citation></ref>
<ref id="B13"><label>13</label><mixed-citation publication-type="journal"><string-name><surname>Coretta</surname>, <given-names>S.</given-names></string-name> (<year>2020</year>). <article-title>Ultrasound Tongue Imaging in R (Version 1.6.0.9000) [R Package]</article-title>.</mixed-citation></ref>
<ref id="B14"><label>14</label><mixed-citation publication-type="journal"><string-name><surname>Davidson</surname>, <given-names>L.</given-names></string-name> (<year>2006</year>). <article-title>Comparing tongue shapes from ultrasound imaging using smoothing spline analysis of variance</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>120</volume>(<issue>1</issue>), <fpage>407</fpage>&#8211;<lpage>415</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.2205133</pub-id></mixed-citation></ref>
<ref id="B15"><label>15</label><mixed-citation publication-type="journal"><string-name><surname>Dawson</surname>, <given-names>K. M.</given-names></string-name>, <string-name><surname>Tiede</surname>, <given-names>M. K.</given-names></string-name>, &amp; <string-name><surname>Whalen</surname>, <given-names>D. H.</given-names></string-name> (<year>2016</year>). <article-title>Methods for quantifying tongue shape and complexity using ultrasound imaging</article-title>. <source>Clinical Linguistics &amp; Phonetics</source>, <volume>30</volume>(<issue>3&#8211;5</issue>), <fpage>328</fpage>&#8211;<lpage>344</lpage>. DOI: <pub-id pub-id-type="doi">10.3109/02699206.2015.1099164</pub-id></mixed-citation></ref>
<ref id="B16"><label>16</label><mixed-citation publication-type="journal"><string-name><surname>Decker</surname>, <given-names>P. M. D.</given-names></string-name>, &amp; <string-name><surname>Nycz</surname>, <given-names>J. R.</given-names></string-name> (<year>2012</year>). <article-title>Are tense [&#230;]s really tense? The mapping between articulation andacoustics</article-title>. <source>Lingua</source>, <volume>122</volume>, <fpage>810</fpage>&#8211;<lpage>821</lpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.lingua.2012.01.003</pub-id></mixed-citation></ref>
<ref id="B17"><label>17</label><mixed-citation publication-type="confproc"><string-name><surname>Diskin</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Loakes</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Billington</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Stoakes</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Gonzalez</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Kirkham</surname>, <given-names>S.</given-names></string-name> (<year>2019</year>). <article-title>The /el/-/&#230;l/ merger in Australian English: Acoustic and articulatory insights</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>1764</fpage>&#8211;<lpage>1768</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B18"><label>18</label><mixed-citation publication-type="webpage"><string-name><surname>Dowle</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Srinivasan</surname>, <given-names>A.</given-names></string-name> (<year>2019</year>). <article-title>data.table: Extension of &#8216;data.frame&#8217; (Version 1.12.2) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=data.table</uri></mixed-citation></ref>
<ref id="B19"><label>19</label><mixed-citation publication-type="journal"><string-name><surname>Gibbon</surname>, <given-names>F. E.</given-names></string-name>, <string-name><surname>Lee</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Yuen</surname>, <given-names>I.</given-names></string-name> (<year>2010</year>). <article-title>Tongue-palate contact during selected vowels in normal speech</article-title>. <source>The Cleft Palate-Craniofacial Journal</source>, <volume>47</volume>(<issue>4</issue>), <fpage>405</fpage>&#8211;<lpage>412</lpage>. DOI: <pub-id pub-id-type="doi">10.1597/09-067.1</pub-id></mixed-citation></ref>
<ref id="B20"><label>20</label><mixed-citation publication-type="journal"><string-name><surname>Gick</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Campbell</surname>, <given-names>F.</given-names></string-name>, &amp; <string-name><surname>Oh</surname>, <given-names>S.</given-names></string-name> (<year>2001</year>). <article-title>A cross-linguistic study of articulatory timing in liquids</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>110</volume>(<issue>5</issue>). DOI: <pub-id pub-id-type="doi">10.1121/1.4777043</pub-id></mixed-citation></ref>
<ref id="B21"><label>21</label><mixed-citation publication-type="thesis"><string-name><surname>Gonzalez</surname>, <given-names>S.</given-names></string-name> (<year>2015</year>). <source>Place oppositions in English coronal obstruents: An ultrasound study</source>. (Doctoral dissertation). <publisher-name>The University of Newcastle</publisher-name>.</mixed-citation></ref>
<ref id="B22"><label>22</label><mixed-citation publication-type="journal"><string-name><surname>Gu</surname>, <given-names>C.</given-names></string-name> (<year>2020</year>). <article-title>General Smoothing Splines (Version 2.1-12) [R Package]</article-title>.</mixed-citation></ref>
<ref id="B23"><label>23</label><mixed-citation publication-type="journal"><string-name><surname>Hewer</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Wuhrer</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Steiner</surname>, <given-names>I.</given-names></string-name>, &amp; <string-name><surname>Richmond</surname>, <given-names>K.</given-names></string-name> (<year>2018</year>). <article-title>A Multilinear Tongue Model Derived from Speech Related MRI Data of the Human Vocal Tract</article-title>. <source>Computer Speech and Language</source>, <volume>51</volume>, <fpage>24</fpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.csl.2018.02.001</pub-id></mixed-citation></ref>
<ref id="B24"><label>24</label><mixed-citation publication-type="journal"><string-name><surname>Heyne</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Derrick</surname>, <given-names>D.</given-names></string-name> (<year>2015</year>). <article-title>Using a radial ultrasound probe&#8217;s virtual origin to compute midsagittal smoothing splines in polar coordinates</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>138</volume>(<issue>6</issue>), <fpage>EL509</fpage>&#8211;<lpage>EL514</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4937168</pub-id></mixed-citation></ref>
<ref id="B25"><label>25</label><mixed-citation publication-type="webpage"><string-name><surname>Hijmans</surname>, <given-names>R. J.</given-names></string-name> (<year>2019</year>). <article-title>raster: Geographic Data Analysis and Modeling (Version 2.9-23) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=raster</uri></mixed-citation></ref>
<ref id="B26"><label>26</label><mixed-citation publication-type="journal"><string-name><surname>Karimi</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Menard</surname>, <given-names>L.</given-names></string-name>, &amp; <string-name><surname>Laporte</surname>, <given-names>C.</given-names></string-name> (<year>2019</year>). <article-title>Fully-automated tongue detection in ultrasound images</article-title>. <source>Computers in Biology and Medicine</source>, <volume>111</volume>, <elocation-id>103335</elocation-id>. DOI: <pub-id pub-id-type="doi">10.1016/j.compbiomed.2019.103335</pub-id></mixed-citation></ref>
<ref id="B27"><label>27</label><mixed-citation publication-type="journal"><string-name><surname>Katz</surname>, <given-names>W. F.</given-names></string-name>, <string-name><surname>Mehta</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Wood</surname>, <given-names>M.</given-names></string-name> (<year>2017</year>). <article-title>Using electromagnetic articulography with a tongue lateral sensor to discriminate manner of articulation</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>141</volume>(<issue>1</issue>), <fpage>7</fpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4973907</pub-id></mixed-citation></ref>
<ref id="B28"><label>28</label><mixed-citation publication-type="journal"><string-name><surname>Kier</surname>, <given-names>W. M.</given-names></string-name>, &amp; <string-name><surname>Smith</surname>, <given-names>K. K.</given-names></string-name> (<year>1985</year>). <article-title>Tongues, tentacles and trunks: The biomechanics of movement in muscular-hydrostats</article-title>. <source>Zoological Journal of the Linnean Society</source>, <volume>83</volume>(<issue>4</issue>). DOI: <pub-id pub-id-type="doi">10.1111/j.1096-3642.1985.tb01178.x</pub-id></mixed-citation></ref>
<ref id="B29"><label>29</label><mixed-citation publication-type="confproc"><string-name><surname>Kocharov</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Evdokimova</surname>, <given-names>V.</given-names></string-name> (<year>2019</year>). <article-title>Within-Word Articulatory Effect of Vowel Rounding</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>552</fpage>&#8211;<lpage>556</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B30"><label>30</label><mixed-citation publication-type="confproc"><string-name><surname>Kochetov</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Faytak</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Nara</surname>, <given-names>K.</given-names></string-name> (<year>2019</year>). <article-title>Manner Differences in The Punjabi Dental-Retroflex Contrast: An Ultrasound Study of Time-Series Data</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>2002</fpage>&#8211;<lpage>2006</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B31"><label>31</label><mixed-citation publication-type="webpage"><string-name><surname>Kunst</surname>, <given-names>J.</given-names></string-name> (<year>2019</year>). <article-title>highcharter: A Wrapper for the &#8216;Highcharts&#8217; Library (Version 0.7.0) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=highcharter</uri></mixed-citation></ref>
<ref id="B32"><label>32</label><mixed-citation publication-type="confproc"><string-name><surname>Lawson</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Stuart-Smith</surname>, <given-names>J.</given-names></string-name> (<year>2019</year>). <article-title>The effects of syllable and sentential position on the timing of lingual gestures in /l/ and /r/</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>547</fpage>&#8211;<lpage>551</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B33"><label>33</label><mixed-citation publication-type="journal"><string-name><surname>Li</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Kambhamettu</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Stone</surname>, <given-names>M.</given-names></string-name> (<year>2005</year>). <article-title>Automatic contour tracking in ultrasound images</article-title>. <source>Clinical Linguistics &amp; Phonetics</source>, <volume>19</volume>(<issue>6&#8211;7</issue>), <fpage>545</fpage>&#8211;<lpage>554</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/02699200500113616</pub-id></mixed-citation></ref>
<ref id="B34"><label>34</label><mixed-citation publication-type="confproc"><string-name><surname>Li</surname>, <given-names>S. R.</given-names></string-name>, <string-name><surname>Woeste</surname>, <given-names>H. M.</given-names></string-name>, <string-name><surname>Dugan</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Mast</surname>, <given-names>T. D.</given-names></string-name>, <string-name><surname>Riley</surname>, <given-names>M. A.</given-names></string-name>, <string-name><given-names>Colin</given-names> <surname>Annand</surname></string-name>, &#8230; <string-name><surname>Boyce</surname>, <given-names>S.</given-names></string-name> (<year>2019</year>). <article-title>Differentiating Normal Vs Misarticulated Tongue Trajectories from Ultrasound for Fast Automatic Articulatory Biofeedback</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>1074</fpage>&#8211;<lpage>1078</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B35"><label>35</label><mixed-citation publication-type="confproc"><string-name><surname>Liker</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Zori&#263;</surname>, <given-names>A. V.</given-names></string-name>, <string-name><surname>Zharkova</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Gibbon</surname>, <given-names>F. E.</given-names></string-name> (<year>2019</year>). <article-title>Ultrasound Analysis of Postalveolar and Palatal Affricates in Croatian: A Case of Neutralisation</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>3666</fpage>&#8211;<lpage>3670</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B36"><label>36</label><mixed-citation publication-type="journal"><string-name><surname>Lim</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Zhu</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Lingala</surname>, <given-names>S. G.</given-names></string-name>, <string-name><surname>Byrd</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Narayanan</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Nayak</surname>, <given-names>K. S.</given-names></string-name> (<year>2019</year>). <article-title>3D dynamic MRI of the vocal tract during natural speech</article-title>. <source>Magn Reson Med</source>, <volume>81</volume>(<issue>3</issue>), <fpage>1511</fpage>&#8211;<lpage>1520</lpage>. DOI: <pub-id pub-id-type="doi">10.1002/mrm.27570</pub-id></mixed-citation></ref>
<ref id="B37"><label>37</label><mixed-citation publication-type="journal"><string-name><surname>Lin</surname>, <given-names>S.</given-names></string-name>, <string-name><surname>Beddor</surname>, <given-names>P. S.</given-names></string-name>, &amp; <string-name><surname>Coetzee</surname>, <given-names>A. W.</given-names></string-name> (<year>2014</year>). <article-title>Gestural reduction, lexical frequency, and sound change: A study of post-vocalic /l/</article-title>. <source>Laboratory Phonology</source>, <volume>5</volume>(<issue>1</issue>), <fpage>9</fpage>&#8211;<lpage>36</lpage>. DOI: <pub-id pub-id-type="doi">10.1515/lp-2014-0002</pub-id></mixed-citation></ref>
<ref id="B38"><label>38</label><mixed-citation publication-type="confproc"><string-name><surname>Maekawa</surname>, <given-names>K.</given-names></string-name> (<year>2019</year>). <article-title>A Real-Time MRI Study of Japanese Moraic Nasal in Utterance-Final Position</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>1987</fpage>&#8211;<lpage>1991</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B39"><label>39</label><mixed-citation publication-type="confproc"><string-name><surname>Mark&#243;</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Bart&#243;k</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Csap&#243;</surname>, <given-names>T. G.</given-names></string-name>, <string-name><surname>Deme</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Gr&#225;czi</surname>, <given-names>T. E.</given-names></string-name> (<year>2019</year>). <article-title>The Effect of Focal Accent on Vowels in Hungarian: Articulatory and Acoustic Data</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name>, &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>2715</fpage>&#8211;<lpage>2719</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B40"><label>40</label><mixed-citation publication-type="journal"><string-name><surname>Mielke</surname>, <given-names>J.</given-names></string-name> (<year>2015</year>). <article-title>An ultrasound study of Canadian French rhotic vowels with polar smoothing spline comparisons</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>137</volume>(<issue>5</issue>). DOI: <pub-id pub-id-type="doi">10.1121/1.4919346</pub-id></mixed-citation></ref>
<ref id="B41"><label>41</label><mixed-citation publication-type="journal"><string-name><surname>Mielke</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Carignan</surname>, <given-names>C.</given-names></string-name>, &amp; <string-name><surname>Thomas</surname>, <given-names>E. R.</given-names></string-name> (<year>2017</year>). <article-title>The articulatory dynamics of pre-velar and pre-nasal /&#230;/-raising in English: An ultrasound study</article-title>. <source>The Journal of the Acoustical Society of America</source>, <volume>142</volume>(<issue>332</issue>), <fpage>332</fpage>&#8211;<lpage>349</lpage>. DOI: <pub-id pub-id-type="doi">10.1121/1.4991348</pub-id></mixed-citation></ref>
<ref id="B42"><label>42</label><mixed-citation publication-type="journal"><string-name><surname>Miller</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Finch</surname>, <given-names>K.</given-names></string-name> (<year>2011</year>). <article-title>Corrected High&#8211;Frame Rate Anchored Ultrasound With Software Alignment</article-title>. <source>Journal of Speech, Language and Hearing Research</source>, <volume>54</volume>, <fpage>16</fpage>. DOI: <pub-id pub-id-type="doi">10.1044/1092-4388(2010/09-0103)</pub-id></mixed-citation></ref>
<ref id="B43"><label>43</label><mixed-citation publication-type="confproc"><string-name><surname>Miller</surname>, <given-names>N. R.</given-names></string-name>, <string-name><surname>Reyes-Aldasoro</surname>, <given-names>C. C.</given-names></string-name>, &amp; <string-name><surname>Verhoeven</surname>, <given-names>J.</given-names></string-name> (<year>2019</year>). <article-title>Asymmetries in Tongue-Palate Contact During Speech</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>1734</fpage>&#8211;<lpage>1738</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B44"><label>44</label><mixed-citation publication-type="confproc"><string-name><surname>Mizoguchi</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Tiede</surname>, <given-names>M. K.</given-names></string-name>, &amp; <string-name><surname>Whalen</surname>, <given-names>D. H.</given-names></string-name> (<year>2019</year>). <article-title>Production of The Japanese Moraic Nasal /N/ by Speakers of English: An Ultrasound Study</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>3493</fpage>&#8211;<lpage>3497</lpage>). <conf-loc>Melbourne, Australia</conf-loc>.</mixed-citation></ref>
<ref id="B45"><label>45</label><mixed-citation publication-type="webpage"><string-name><surname>Neuwirth</surname>, <given-names>E.</given-names></string-name> (<year>2014</year>). <article-title>RColorBrewer: ColorBrewer Palettes (Version 1.1-2) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=RColorBrewer</uri></mixed-citation></ref>
<ref id="B46"><label>46</label><mixed-citation publication-type="confproc"><string-name><surname>Oakley</surname>, <given-names>M.</given-names></string-name> (<year>2019</year>). <article-title>Articulation of L2 French Mid and High Vowels</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>1724</fpage>&#8211;<lpage>1728</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B47"><label>47</label><mixed-citation publication-type="webpage"><string-name><surname>Owen</surname>, <given-names>J.</given-names></string-name> (<year>2018</year>). <article-title>rhandsontable: Interface to the &#8216;Handsontable.js&#8217; Library (Version 0.3.7) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=rhandsontable</uri></mixed-citation></ref>
<ref id="B48"><label>48</label><mixed-citation publication-type="webpage"><string-name><surname>Pebesma</surname>, <given-names>E. J.</given-names></string-name>, &amp; <string-name><surname>Bivand</surname>, <given-names>R. S.</given-names></string-name> (<year>2005</year>). <article-title>Classes and methods for spatial data in R [R package]</article-title>. Retrieved from <uri>https://cran.r-project.org/doc/Rnews/</uri></mixed-citation></ref>
<ref id="B49"><label>49</label><mixed-citation publication-type="journal"><string-name><surname>Pini</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Spreafico</surname>, <given-names>L.</given-names></string-name>, <string-name><surname>Vantini</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Vietti</surname>, <given-names>A.</given-names></string-name> (<year>2019</year>). <article-title>Multi-aspect local inference for functional data: Analysis of ultrasound tongue profiles</article-title>. <source>Journal of Multivariate Analysis</source>, <volume>170</volume>, <fpage>13</fpage>. DOI: <pub-id pub-id-type="doi">10.1016/j.jmva.2018.11.006</pub-id></mixed-citation></ref>
<ref id="B50"><label>50</label><mixed-citation publication-type="thesis"><string-name><surname>Proctor</surname>, <given-names>M.</given-names></string-name> (<year>2009</year>). <source>Gestural Characterization of a Phonological Class: The Liquids</source>. (Doctoral dissertation). <publisher-name>Yale University</publisher-name>.</mixed-citation></ref>
<ref id="B51"><label>51</label><mixed-citation publication-type="confproc"><string-name><surname>Proctor</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Lo</surname>, <given-names>C. Y.</given-names></string-name>, &amp; <string-name><surname>Narayanan</surname>, <given-names>S.</given-names></string-name> (<year>2015</year>). <source>Articulation of English vowels in running speech: A real-time MRI study</source>. <conf-name>Paper presented at the 18th International Congress of Phonetic Sciences</conf-name>, <conf-loc>London</conf-loc>.</mixed-citation></ref>
<ref id="B52"><label>52</label><mixed-citation publication-type="webpage"><collab>R Core Team</collab>. (<year>2018</year>). <article-title>R: A Language and Environment for Statistical Computing: R Foundation for Statistical Computing</article-title>. Retrieved from <uri>https://www.R-project.org/</uri></mixed-citation></ref>
<ref id="B53"><label>53</label><mixed-citation publication-type="webpage"><string-name><surname>Ram</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Wickham</surname>, <given-names>H.</given-names></string-name> (<year>2018</year>). <article-title>wesanderson: A Wes Anderson Palette Generator (Version 0.3.6.9000) [R package]</article-title>. Retrieved from <uri>https://github.com/karthik/wesanderson</uri></mixed-citation></ref>
<ref id="B54"><label>54</label><mixed-citation publication-type="confproc"><string-name><surname>Recasens</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Rodri&#769;guez</surname>, <given-names>C.</given-names></string-name> (<year>2019</year>). <article-title>The Effect of Voicing on Tongue Configuration for Unaspirated Stop Sequences</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>418</fpage>&#8211;<lpage>421</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B55"><label>55</label><mixed-citation publication-type="confproc"><string-name><surname>Roon</surname>, <given-names>K. D.</given-names></string-name>, &amp; <string-name><surname>Whalen</surname>, <given-names>D. H.</given-names></string-name> (<year>2019</year>). <article-title>Velarization of Russian labial consonants</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>3488</fpage>&#8211;<lpage>3492</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B56"><label>56</label><mixed-citation publication-type="confproc"><string-name><surname>Shadle</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Proctor</surname>, <given-names>M. I.</given-names></string-name>, &amp; <string-name><surname>Iskarous</surname>, <given-names>K.</given-names></string-name> (<year>2008</year>). <source>An MRI Study of the Effect of Vowel Context on English Fricatives</source>. <conf-name>Paper presented at the Joint Meeting of the Acoustical Society of America and European Acoustics Association</conf-name>, <conf-loc>Paris, France</conf-loc>. DOI: <pub-id pub-id-type="doi">10.1121/1.2935246</pub-id></mixed-citation></ref>
<ref id="B57"><label>57</label><mixed-citation publication-type="webpage"><string-name><surname>Sievert</surname>, <given-names>C.</given-names></string-name> (<year>2018</year>). <article-title>plotly for R</article-title>. Retrieved from <uri>https://plotly-r.com</uri></mixed-citation></ref>
<ref id="B58"><label>58</label><mixed-citation publication-type="webpage"><string-name><surname>Slowikowski</surname>, <given-names>K.</given-names></string-name> (<year>2019</year>). <article-title>ggrepel: Automatically Position Non-Overlapping Text Labels with &#8216;ggplot2&#8217; (Version 0.8.1) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=ggrepel</uri></mixed-citation></ref>
<ref id="B59"><label>59</label><mixed-citation publication-type="journal"><string-name><surname>Stone</surname>, <given-names>M.</given-names></string-name> (<year>2005</year>). <article-title>A guide to analysing tongue motion from ultrasound images</article-title>. <source>Clinical Linguistics &amp; Phonetics</source>, <volume>19</volume>(<issue>6&#8211;7</issue>), <fpage>455</fpage>&#8211;<lpage>501</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/02699200500113558</pub-id></mixed-citation></ref>
<ref id="B60"><label>60</label><mixed-citation publication-type="journal"><string-name><surname>Stone</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Goldstein</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Zhang</surname>, <given-names>Y.</given-names></string-name> (<year>1997</year>). <article-title>Principal component analysis of cross sections of tongue shapes in vowel production</article-title>. <source>Speech Communication</source>, <volume>22</volume>(<issue>2&#8211;3</issue>), <fpage>11</fpage>. DOI: <pub-id pub-id-type="doi">10.1016/S0167-6393(97)00027-7</pub-id></mixed-citation></ref>
<ref id="B61"><label>61</label><mixed-citation publication-type="confproc"><string-name><surname>Stone</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Murano</surname>, <given-names>E. Z.</given-names></string-name> (<year>2007</year>). <source>Speech patterns in a muscular hydrostat: Lip, tongue and glossectomy movement</source>. <conf-name>Paper presented at the Proceedings of the Third B-J-K Symposium on Biomechanics, Healthcare and Information Science</conf-name>, <conf-loc>Kanazawa, Japan</conf-loc>.</mixed-citation></ref>
<ref id="B62"><label>62</label><mixed-citation publication-type="confproc"><string-name><surname>Strycharczuk</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Scobbie</surname>, <given-names>J. M.</given-names></string-name> (<year>2015</year>). <source>Velocity measures in ultrasound data. Gestural timing of post-vocalic /l/ in English</source>. <conf-name>Paper presented at the Proceedings of the 18th International Congress on Phonetic Sciences</conf-name>.</mixed-citation></ref>
<ref id="B63"><label>63</label><mixed-citation publication-type="webpage"><string-name><surname>Tang</surname>, <given-names>Y.</given-names></string-name> (<year>2019</year>). <article-title>shinyjqui: &#8216;jQuery UI&#8217; Interactions and Effects for Shiny (Version 0.3.2) [R package]</article-title>. Retrieved from <uri>https://github.com/yang-tang/shinyjqui</uri></mixed-citation></ref>
<ref id="B64"><label>64</label><mixed-citation publication-type="webpage"><string-name><surname>Thieurmel</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Marcelionis</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Petit</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Salette</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Robert</surname>, <given-names>T.</given-names></string-name> (<year>2019</year>). <article-title>rAmCharts: JavaScript Charts Tool (Version 2.1.10) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=rAmCharts</uri></mixed-citation></ref>
<ref id="B65"><label>65</label><mixed-citation publication-type="journal"><string-name><surname>Verhoeven</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Miller</surname>, <given-names>N. R.</given-names></string-name>, <string-name><surname>Daems</surname>, <given-names>L.</given-names></string-name>, &amp; <string-name><surname>Reyes-Aldasoro</surname>, <given-names>C. C.</given-names></string-name> (<year>2019</year>). <article-title>Visualisation and Analysis of Speech Production with Electropalatography</article-title>. <source>Journal of Imaging</source>, <volume>5</volume>(<issue>40</issue>), <fpage>16</fpage>. DOI: <pub-id pub-id-type="doi">10.3390/jimaging5030040</pub-id></mixed-citation></ref>
<ref id="B66"><label>66</label><mixed-citation publication-type="journal"><string-name><surname>Verma</surname>, <given-names>S. K.</given-names></string-name>, <string-name><surname>Tandon</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Agrawal</surname>, <given-names>D. K.</given-names></string-name>, &amp; <string-name><surname>Prabhat</surname>, <given-names>K. C.</given-names></string-name> (<year>2012</year>). <article-title>A cephalometric evaluation of tongue from the rest position to centric occlusion in the subjects with class II division 1 malocclusion and class I normal occlusion</article-title>. <source>Journal of Orthodontic Science</source>, <volume>1</volume>(<issue>2</issue>), <fpage>34</fpage>&#8211;<lpage>39</lpage>. DOI: <pub-id pub-id-type="doi">10.4103/2278-0203.99758</pub-id></mixed-citation></ref>
<ref id="B67"><label>67</label><mixed-citation publication-type="book"><string-name><surname>Wickham</surname>, <given-names>H.</given-names></string-name> (<year>2016</year>). <source>ggplot2: Elegant Graphics for Data Analysis</source>. <publisher-loc>New York</publisher-loc>: <publisher-name>Springer-Verlag</publisher-name>. DOI: <pub-id pub-id-type="doi">10.1007/978-3-319-24277-4</pub-id></mixed-citation></ref>
<ref id="B68"><label>68</label><mixed-citation publication-type="webpage"><string-name><surname>Wickham</surname>, <given-names>H.</given-names></string-name> (<year>2018</year>). <article-title>pryr: Tools for Computing on the Language (Version 0.1.4) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=pryr</uri></mixed-citation></ref>
<ref id="B69"><label>69</label><mixed-citation publication-type="webpage"><string-name><surname>Wickham</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>Henry</surname>, <given-names>L.</given-names></string-name> (<year>2019</year>). <article-title>tidyr: Easily Tidy Data with &#8216;spread()&#8217; and &#8216;gather()&#8217; Functions (Version 0.8.3) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=tidyr</uri></mixed-citation></ref>
<ref id="B70"><label>70</label><mixed-citation publication-type="book"><string-name><surname>Wrench</surname>, <given-names>A.</given-names></string-name> (<year>2012</year>). <chapter-title>Articulate Assistant Advanced User Guide</chapter-title>. <publisher-loc>Edinburgh</publisher-loc>: <publisher-name>Articulate Instruments Ltd</publisher-name>.</mixed-citation></ref>
<ref id="B71"><label>71</label><mixed-citation publication-type="webpage"><string-name><surname>Xie</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Cheng</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Tan</surname>, <given-names>X.</given-names></string-name> (<year>2019</year>). <article-title>DT: A Wrapper of the JavaScript Library &#8216;DataTables&#8217; (Version 0.8) [R package]</article-title>. Retrieved from <uri>https://CRAN.R-project.org/package=DT</uri></mixed-citation></ref>
<ref id="B72"><label>72</label><mixed-citation publication-type="confproc"><string-name><surname>Zeroual</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Hoole</surname>, <given-names>P.</given-names></string-name>, &amp; <string-name><surname>Gafos</surname>, <given-names>A.</given-names></string-name> (<year>2019</year>). <article-title>Vowel-To-Consonant Coarticulation In Moroccan Arabic</article-title>. In <string-name><given-names>S.</given-names> <surname>Calhoun</surname></string-name>, <string-name><given-names>P.</given-names> <surname>Escudero</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Tabain</surname></string-name> &amp; <string-name><given-names>P.</given-names> <surname>Warren</surname></string-name> (Eds.), <conf-name>Proceedings of the 19th International Congress of Phonetic Sciences</conf-name> (pp. <fpage>537</fpage>&#8211;<lpage>541</lpage>). <conf-loc>Melbourne, Australia</conf-loc>: <conf-sponsor>Australasian Speech Science and Technology Association Inc</conf-sponsor>.</mixed-citation></ref>
<ref id="B73"><label>73</label><mixed-citation publication-type="journal"><string-name><surname>Zharkova</surname>, <given-names>N.</given-names></string-name> (<year>2013</year>). <article-title>Using ultrasound to quantify tongue shape and movement characteristics</article-title>. <source>The Cleft Palate-Craniofacial Journal</source>, <volume>50</volume>(<issue>1</issue>), <fpage>76</fpage>&#8211;<lpage>81</lpage>. DOI: <pub-id pub-id-type="doi">10.1597/11-196</pub-id></mixed-citation></ref>
<ref id="B74"><label>74</label><mixed-citation publication-type="journal"><string-name><surname>Zharkova</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Gibbon</surname>, <given-names>F. E.</given-names></string-name>, &amp; <string-name><surname>Lee</surname>, <given-names>A.</given-names></string-name> (<year>2017</year>). <article-title>Using ultrasound tongue imaging to identify covert contrasts in children&#8217;s speech</article-title>. <source>Clinical Linguistics &amp; Phonetics</source>, <volume>31</volume>(<issue>1</issue>), <fpage>21</fpage>&#8211;<lpage>34</lpage>. DOI: <pub-id pub-id-type="doi">10.1080/02699206.2016.1180713</pub-id></mixed-citation></ref>
</ref-list>
</back>
</article>