<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.2 20120330//EN" "http://jats.nlm.nih.gov/publishing/1.2/JATS-journalpublishing1.dtd">
<!--<?xml-stylesheet type="text/xsl" href="article.xsl"?>-->
<article article-type="research-article" dtd-version="1.2" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance">
<front>
<journal-meta>
<journal-id journal-id-type="issn">1868-6354</journal-id>
<journal-title-group>
<journal-title>Laboratory Phonology: Journal of the Association for Laboratory Phonology</journal-title>
</journal-title-group>
<issn pub-type="epub">1868-6354</issn>
<publisher>
<publisher-name>Open Library of Humanities</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="doi">10.16995/labphon.16542</article-id>
<article-categories>
<subj-group>
<subject>Journal article</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Learning new speech sounds in remote and in-person protocols: Benefits, drawbacks, and considerations for future research</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Baese-Berk</surname>
<given-names>Melissa M.</given-names>
</name>
<email>mmbb@uchicago.edu</email>
<xref ref-type="aff" rid="aff-1">1</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Staggs</surname>
<given-names>Cecelia</given-names>
</name>
<email>cecelias@uoregon.edu</email>
<xref ref-type="aff" rid="aff-2">2</xref>
</contrib>
<contrib contrib-type="author">
<name>
<surname>Jaramillo</surname>
<given-names>Santiago</given-names>
</name>
<email>sjara@uoregon.edu</email>
<xref ref-type="aff" rid="aff-3">3</xref>
</contrib>
</contrib-group>
<aff id="aff-1"><label>1</label>Department of Linguistics, University of Chicago, Chicago, IL, USA</aff>
<aff id="aff-2"><label>2</label>Department of Linguistics, University of Oregon, Eugene, OR, USA</aff>
<aff id="aff-3"><label>3</label>Institute of Neuroscience, University of Oregon, Eugene, OR, USA</aff>
<pub-date publication-format="electronic" date-type="pub" iso-8601-date="2025-05-16">
<day>16</day>
<month>05</month>
<year>2025</year>
</pub-date>
<pub-date pub-type="collection">
<year>2025</year>
</pub-date>
<volume>16</volume>
<issue>1</issue>
<fpage>1</fpage>
<lpage>25</lpage>
<permissions>
<copyright-statement>Copyright: &#x00A9; 2025 The Author(s)</copyright-statement>
<copyright-year>2025</copyright-year>
<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International License (CC-BY 4.0), which permits unrestricted use, distribution, and reproduction in any medium, provided the original author and source are credited. See <uri xlink:href="http://creativecommons.org/licenses/by/4.0/">http://creativecommons.org/licenses/by/4.0/</uri>.</license-p>
</license>
</permissions>
<self-uri xlink:href="http://www.journal-labphon.org/articles/10.16995/labphon.16542/"/>
<abstract>
<p>In-laboratory training of novel speech sounds has provided significant insight into how adult language learners learn new sounds. However, this training is often costly in terms of time in lab for participants and for experimenters. Therefore, understanding whether such paradigms can be conducted successfully in a remote setting has been of great interest to the field. In this study, we present data from participants in both in-lab and remote protocols for a two-day language learning paradigm. We compare both the results of learning across the two paradigms and the logistical aspects of conducting this research. We demonstrate that both paradigms result in learning; however, while data collection is much faster for remote protocols, there are significant trade-offs to consider in terms of individual performance and attrition within this population, among other concerns. We discuss these concerns from a methodological standpoint and explore how comparisons of training modality can be informative not only for choosing appropriate methodology but also for developing a deeper understanding of acquisition of novel speech sounds.</p>
</abstract>
</article-meta>
</front>
<body>
<sec>
<title>1. Introduction</title>
<p>For decades, researchers have used in-lab training protocols to investigate how individuals learn new speech sounds in language (e.g., <xref ref-type="bibr" rid="B46">Lively et al., 1993</xref>; <xref ref-type="bibr" rid="B47">Logan et al., 1991</xref>; <xref ref-type="bibr" rid="B68">Strange &amp; Dittman, 1984</xref>). This work has been extremely informative and has led to insights in methodological aspects of training that impact learning (e.g., high variability phonetic training, <xref ref-type="bibr" rid="B13">Bradlow et al., 1997</xref>, <xref ref-type="bibr" rid="B12">1999</xref> <xref ref-type="bibr" rid="B47">Logan et al., 1991</xref>), cognitive factors that impact learning (e.g., <xref ref-type="bibr" rid="B57">Perrachione et al., 2011</xref>), and the consequences for linguistic systems more broadly after learning (e.g., <xref ref-type="bibr" rid="B39">Kartushina et al., 2016</xref>). However, this training is also extremely costly. Typical training paradigms take place over multiple days, with some studies taking as long as 45 sessions in the lab. Indeed, multiple day training has been shown to be highly effective and perhaps even necessary for this type of learning to occur (e.g., <xref ref-type="bibr" rid="B26">Earle &amp; Myers, 2015</xref>).</p>
<p>Given this extended time requirement for data collection, such training studies are often difficult to run. Bringing participants into a lab for multiple days has logistical challenges in terms of scheduling, and these training paradigms often require significant experimenter oversight, requiring a lab to be staffed with a number of researchers who can conduct this training.</p>
<p>Even before the COVID-19 crisis, there was significant interest in conducting psycholinguistic and laboratory phonology studies remotely for many reasons. Prime among them is ease of recruitment and ease of conducting the research (<xref ref-type="bibr" rid="B63">Schnoebelen &amp; Kuperman, 2010</xref>). However, conducting data collection remotely has other benefits, including the ability to recruit a more diverse population of participants (<xref ref-type="bibr" rid="B56">Pavlick et al., 2014</xref>).</p>
<p>Given these potential benefits, it is tempting to explore options to conduct training studies fully remotely. However, it is not yet known how performance on these tasks may differ in remote and in-lab protocols. For example, because attention has been shown as a critical component of laboratory-based training protocols (e.g., <xref ref-type="bibr" rid="B54">Mora &amp; Mora-Plaza, 2019</xref>), it is possible that participants completing remote protocols have too many distractions in their environment to result in the same type of learning that we typically see in the laboratory. Therefore, it is critical to compare in-laboratory training to remote data collection to ensure the two are comparable.</p>
<p>Below, we present a summary of previous findings from in-laboratory training for novel speech sounds, focusing especially on areas where we believe in-laboratory and remote data collection may differ from one another. We then provide a summary of remote data collection in laboratory phonology, focusing first on data collected before March of 2020 (the onset of the COVID-19 crisis in the United States) and then on data collected remotely during the pandemic phase.</p>
<sec>
<title>1.1. In-laboratory training for novel speech sounds</title>
<p>In-laboratory training has been used to demonstrate acquisition of very difficult sound contrasts in a learner&#8217;s second language (L2). For example, in a series of landmark papers, Logan and colleagues examined learning of English /&#633;/ and /l/ by native Japanese speakers (<xref ref-type="bibr" rid="B13">Bradlow et al., 1997</xref>, <xref ref-type="bibr" rid="B12">1999</xref> <xref ref-type="bibr" rid="B46">Lively et al., 1993</xref>; <xref ref-type="bibr" rid="B47">Logan et al., 1991</xref>). In these studies, they specifically examined how learners both perceive and produce the /&#633;/ /l/ contrast after significant amounts of training. This work was particularly groundbreaking because non-native contrasts had previously been thought to be extremely challenging to acquire in adulthood, and perhaps impervious to learning, even after significant exposure. Further, there was a concern that decontextualized learning in the laboratory may be inferior to classroom training and thus may not be amenable to investigating the processes underlying learning of novel speech sounds. In this set of studies, they not only demonstrated robust learning, but also elucidated some possible individual factors that impact learning. This work was critical for demonstrating that learning of novel speech sounds is possible in the laboratory, and that the laboratory is a valid testing ground for investigating myriad factors that may impact learning. Since these groundbreaking studies, researchers have investigated various aspects of training, including the amount of variability presented to participants during training (<xref ref-type="bibr" rid="B37">Iverson et al., 2005</xref>; <xref ref-type="bibr" rid="B47">Logan et al., 1991</xref>). In general, these aspects of training are a property of the stimuli and could remain stable in remote data collection, therefore we do not consider these studies further here.</p>
<p>Similarly, many linguistic factors have been shown to impact learning in these paradigms. Specifically, the relationship between a first and second language&#8217;s phonological system has been demonstrated to be critically important to the ease or difficulty of learning new speech sounds (<xref ref-type="bibr" rid="B11">Best &amp; Tyler, 2007</xref>; <xref ref-type="bibr" rid="B30">Flege &amp; Bohn, 2021</xref>). While these factors are also relatively easy to control in remote data collection, it is important to note that recruitment in remote protocols often requires participants to self-report about their language background. Many in-laboratory studies also use self-reporting for language background information, but it may be more difficult to verify this information in remote protocols that typically do not have a live experimenter observing the session. That is, it is possible a researcher in an in-person study could flag a participant as possibly not meeting the language background requirements, and this may not be possible in remote settings. Further, in remote protocols, participants have been shown to give incorrect information about themselves to access more studies (e.g., <xref ref-type="bibr" rid="B18">Chandler &amp; Paolacci, 2017</xref>; <xref ref-type="bibr" rid="B1">Aguinis et al., 2021</xref>).</p>
<p>Other aspects that have been shown to influence learning of novel speech sounds include a variety of cognitive factors, which are typically described under the umbrella of &#8220;individual differences.&#8221; These properties are typically thought of as being relatively static. However, it is also known that some of these properties are impacted by external factors. For example, working memory has been shown to impact learning of L2 sounds (<xref ref-type="bibr" rid="B24">Darcy et al., 2015</xref>). However, working memory is not a static property within an individual, as it is impacted by distraction (e.g., <xref ref-type="bibr" rid="B70">West, 1999</xref>). Further, working memory and stress states have a reciprocal relationship (e.g., <xref ref-type="bibr" rid="B49">Matthews &amp; Campbell, 2009</xref>), suggesting that working memory may vary substantially across testing environments.</p>
<p>Similarly, it is relatively easy to control for issues of attention and cognitive load during in-laboratory studies. In general, the circumstances for all participants in a given in-laboratory study are similar. That is, all participants in a particular study are usually tested in the same (or very similar) environments with limited distractions and are presumed to be primarily attending to the target task. However, this is not something that can be controlled in remote environments. While one participant may complete the experiment in a quiet, private room, others may complete the experiment in a noisy (both auditorily and visually) public space where there are competing demands on their attention. Because attentional control has been demonstrated to impact L2 speech sound learning (e.g., <xref ref-type="bibr" rid="B53">Mora &amp; Darcy, 2023</xref>), it is possible these differences in attention and distractions during testing could impact participants differently in ways that are difficult to assess or predict in remote testing environments. While cognitive load has not been studied in as much detail in L2 speech sound learning (cf. <xref ref-type="bibr" rid="B7">Baese-Berk &amp; Samuel, 2016</xref>), it is possible that distractions in the environment could impact a learners&#8217; ability to acquire novel speech sounds in remote testing environments by increasing their cognitive load during learning.</p>
<p>In-laboratory training and testing is often limited in terms of the participants available to a researcher. That is, in much of the research in this area, investigations are limited to populations which are geographically close to an experimenter&#8217;s laboratory. This has resulted in samples from &#8220;WEIRD&#8221; societies (Western, Educated, Industrialized, Rich and Democratic) (<xref ref-type="bibr" rid="B33">Henrich et al., 2010</xref>; see <xref ref-type="bibr" rid="B55">Ortega, 2005</xref>, and <xref ref-type="bibr" rid="B3">Andringa &amp; Godfroid, 2020</xref>, for a review of this issue in the applied linguistics domain). Indeed, in many L2 speech sound learning experiments, the typical participant for an in-laboratory study is a college student from a WEIRD society, resulting in participants who are even less representative of the general population than if participants were sampled broadly from the same community. Further, it is likely that in-laboratory studies that are not part of a university course also select for a specific type of person who is, perhaps, especially interested in the research topic, organized enough to set up times to come into the lab, or may have other special personality features. However, these features may not be common among everyone who wants to (or needs to) learn additional languages.<xref ref-type="fn" rid="n1">1</xref> Therefore, in-laboratory work has resulted in findings that are perhaps not robustly generalizable to a broader population.</p>
<p>In the present study, we treat the in-laboratory participants as the baseline and compare the remote group to them. This is in part because, as a field, we feel as though we have a better sense of what in-laboratory participants are like and what they can and cannot do as compared to remote participants. This is likely because we have decades of work with in-lab participants and relatively less experience with remote participants. However, it is possible that the in-lab participants are not a sufficient baseline for comparison. That is, in-person methods are likely to create their own non-representative patterns or artifacts, especially because the laboratory context deviates significantly from a participants&#8217; daily experience. This has been demonstrated in previous work that has shown that being in a laboratory is sufficient to prime performance on a task using altered auditory feedback (e.g., <xref ref-type="bibr" rid="B36">Houde &amp; Jordan, 2002</xref>).</p>
</sec>
<sec>
<title>1.2. Remote data collection in laboratory phonology</title>
<p>Because we are unaware of any work that directly compares remote data collection for learning novel speech sounds to in-laboratory learning, we address here the basic issues of remote data collection for laboratory phonology and psycholinguistic research more broadly, which underlie the type of training we conduct in the present study.</p>
<sec>
<title>1.2.1. Pre-pandemic</title>
<p>Even before the COVID-19 pandemic necessitated the cessation of in-laboratory data collection and spurred many researchers to consider internet-based, remote data collection, researchers in linguistics, cognitive science, and psychology had already begun to make moves toward remote data collection to supplement or, in some cases, replace in-laboratory data collection. Many researchers championed the convenience of online data collection, especially in cases where in-person work was challenging. For example, Kimball and colleagues (<xref ref-type="bibr" rid="B41">2019</xref>) discussed ways to expand field studies and work with under-resourced languages using remote data collection. In other cases, recruitment of specific populations of participants within a community is challenging and thus remote data collection improves researchers&#8217; ability to investigate populations often not well-represented among college students (<xref ref-type="bibr" rid="B65">Staggs et al., 2022</xref>). It is also the case that some institutions and researchers do not have appropriate resources for in person data collection; therefore, remote data collection can democratize scientific investigation (<xref ref-type="bibr" rid="B40">Kimball, 2014</xref>).</p>
<p>However, even outside the recruitment of specific populations or the use of remote data collection to alleviate other challenging research environments, researchers in behavioral sciences have used online and remote data collection methods for many years. Many of these earlier studies focused on the feasibility of using online data collection tools, including Amazon&#8217;s Mechanical Turk, Prolific, and FindingFive (<xref ref-type="bibr" rid="B16">Buhrmester et al., 2011</xref>; <xref ref-type="bibr" rid="B23">Crump et al., 2013</xref>; <xref ref-type="bibr" rid="B48">Mason &amp; Suri, 2012</xref>), to conduct psycholinguistic research and replicate well-known in-laboratory effects. Indeed, many speech scientists began to use these tools to collect data and demonstrated that performance in these remote protocols was similar to in-laboratory experiments (e.g., <xref ref-type="bibr" rid="B21">Chodroff &amp; Wilson, 2014</xref>; <xref ref-type="bibr" rid="B72">Yu &amp; Lee, 2014</xref>). These tools were used for collecting perceptual judgments of speech (e.g., <xref ref-type="bibr" rid="B44">Kunath &amp; Weinberger, 2010</xref>) and reaction time data (e.g., <xref ref-type="bibr" rid="B29">Enochson &amp; Culbertson, 2015</xref>), among other measures.</p>
<p>The consensus before the pandemic was that these tools were extremely useful for data collection, as the typically onerous process of bringing participants into the lab to conduct studies could be made much easier by collecting data remotely. Researchers also noted other benefits, including recruiting participants from a more diverse population than what is typically available in a university human subject pool.</p>
</sec>
<sec>
<title>1.2.2. During the pandemic</title>
<p>Of course, at the height of the pandemic, interest in remote data collection expanded exponentially. Many (remote) conferences held special sessions or tutorials to demonstrate how to implement remote data collection (e.g., Rachel Theodore at the 2021 meeting of the Acoustical Society of America). With this proliferation of new studies, an interest in more closely comparing in person and remote data collection also grew. In an introduction to a special issue of <italic>Linguistic Vanguard</italic> around how remote data collection can be used effectively, Kostadinova and Gardner (<xref ref-type="bibr" rid="B43">2024</xref>) note that remote data collection often yields data on par with in-laboratory data collection. While many effects were replicated when data was collected remotely, some studies failed to replicate well-known results (<xref ref-type="bibr" rid="B14">Brekelmans et al., 2022</xref>; <xref ref-type="bibr" rid="B31">Grieve, 2021</xref>), leading to some skepticism from researchers and reviewers about the validity of online data collection. Further, remote vs. in-lab data collection interacts in important ways with task, suggesting that not all tasks are equally appropriate to be conducted online (<xref ref-type="bibr" rid="B15">Bro&#347;, 2025</xref>).</p>
<p>Questions about why such differences emerge has been a topic of substantial debate. For example, a recent paper demonstrated that participants tested in remote setups performed better on some tasks than participants completing tasks with an experimenter present (<xref ref-type="bibr" rid="B10">Bent et al., 2024</xref>). The authors suggest that these findings may be driven by the presence of an experimenter, which could shift participant behavior in a negative way or by demographic characteristics of the participants, which differed across the two data collection situations. Understanding whether and why differences might emerge across means of data collection is critically important if one hopes to use remote data collection to supplement or replace in-laboratory studies. Therefore, it is necessary to compare performance and logistical aspects of in-laboratory and remote data collection across a variety of tasks.</p>
</sec>
</sec>
<sec>
<title>1.3. Current study</title>
<p>In this study, we compare in-laboratory and remotely collected data for training of discrimination of novel speech sounds. As described above, these types of training studies are often quite challenging to conduct in person because they typically require multiple days of training which have logistical and scheduling challenges for both participants and researchers. Therefore, if remote data collection yields similar results as in-laboratory data collection, one could feasibly replace or supplement data collection using logistically simpler procedures.</p>
<p>Below, we describe a two-day training study which was conducted both in-laboratory and using remote data collection methods. We present the results and discuss both methodological and theoretical implications of these results.</p>
</sec>
</sec>
<sec>
<title>2. Methods</title>
<sec>
<title>2.1. Participants</title>
<p>Participants for the in-laboratory condition of the experiment were recruited via the Linguistics and Psychology Human Subject Pool at the University of Oregon. We set a one-term window for recruitment, so all participants were recruited during a single 10-week quarter. Eighteen participants completed one day of the in-laboratory experiment and 16 of those participants completed both days of training.</p>
<p>Participants for the remote condition of the experiment were recruited via Prolific. We set a 30-participant cap for the experiment, and full completion of this condition of the study took two days. On day one, all participants were recruited and completed the first day of training, but not everyone took part in the training on day two. In total, 30 participants completed one day of training, and 22 participants completed both days of training.</p>
<p>All participants reported their ages as being between 18&#8211;35 years (mean = 21.6). They reported that their first language was English, and they were not proficient in any other languages. Further, no participants reported experience with any languages that use a dental-retroflex contrast (e.g., Hindi), which was used in this study (see Section 2.2, below). Participants in both conditions were paid for their participation.</p>
</sec>
<sec>
<title>2.2. Materials</title>
<p>Stimuli were drawn from a synthetic continuum originally created by Stevens and Blumstein (<xref ref-type="bibr" rid="B66">1975</xref>). They consist of a six-step continuum from / &#598;a/ to / d&#810;a/, representing the Hindi dental-retroflex contrast. They vary in the onset of the second and third formant frequencies and in the frequency of the burst. Specific details about the creation of these stimuli can be found in Stevens and Blumstein. All stimuli were normed for duration and were amplitude normalized.</p>
</sec>
<sec>
<title>2.3. Procedure</title>
<p>The procedures for the in-laboratory and remote experiments were designed to be identical, except for location of the experiment. On the first day of training, participants began the experiment by giving informed consent for participation and completing a brief language background questionnaire to ensure they were members of our target demographic age range and language experience. Following this portion of the experiment, participants in the remote condition completed a headphone check (adapted from <xref ref-type="bibr" rid="B71">Woods et al., 2017</xref>) to ensure they were completing the experiment using headphones.<xref ref-type="fn" rid="n2">2</xref> No participants were excluded based on the results of the headphone check.</p>
<p>Each of two days of training then followed the same pattern: Participants first completed a pre-test, then training, and then the post-test. During all portions of the experiment, participants completed an ABX discrimination task. They heard two different sounds from the continuum (&#8220;A&#8221; and &#8220;B&#8221;) and then heard a third sound (&#8220;X&#8221;), which was identical to either &#8220;A&#8221; or &#8220;B&#8221;. They were asked to identify which of the two sounds they heard by pressing either &#8220;f&#8221; or &#8220;j&#8221; on their keyboard.<xref ref-type="fn" rid="n3">3</xref> The pre-and post-tests consisted of 72 trials each (a total of 288 trials across the pre- and post-tests on the two days). The training portion consisted of 504 trials per day (a total of 1008 trials across two days). This training and testing protocol is identical to those used in previous studies (e.g., <xref ref-type="bibr" rid="B7">Baese-Berk &amp; Samuel, 2016</xref>, <xref ref-type="bibr" rid="B5">2022</xref>).</p>
</sec>
<sec>
<title>2.4. Analysis</title>
<p>Data were analyzed following procedures in Baese-Berk and Samuel (<xref ref-type="bibr" rid="B7">2016</xref>, <xref ref-type="bibr" rid="B8">2022</xref>) using logistic mixed effects regressions to estimate the proportion correct for discrimination of each pair of stimuli. The dependent variable is whether the participants responded correctly (coded as &#8220;1&#8221;) or incorrectly (coded as &#8220;0&#8221;) to the discrimination task.</p>
<p>The full model included a comparison of all the peripheral pairs to the center pair (i.e., coded so that the average of the other pairs was compared to Pair 3&#8211;4, with the non-central pairs as reference level), training modality (In-laboratory reference level vs. Remote), training phase (Day 1 Pre-Test reference level vs. Day 2 Post-Test), and interactions between these factors as well as a random intercept for participants.</p>
<p>Given previous work in categorical perception, the expectation was that participants should perform at chance for discrimination among all adjacent pairs at pre-test. At post-test, if participants have learned the novel speech sound contrast, they should perform at chance for pairs away from the category boundary (i.e., pairs 1&#8211;2, 2&#8211;3, 4&#8211;5, and 5&#8211;6) and above chance for the pair that crosses the category boundary (i.e., pair 3&#8211;4).<xref ref-type="fn" rid="n4">4</xref></p>
</sec>
</sec>
<sec>
<title>3. Results</title>
<p>Below, we present the results of learning for each training condition separately, followed by a comparison of the two modalities. We begin with a general description of the trends in the data and follow this with the statistical analysis described above. Next, we present a discussion of individual performance and variability in this performance, and a discussion of the logistical aspects faced during completion of the experiments.</p>
<sec>
<title>3.1. Learning after in-lab training</title>
<p>At pre-test, in-laboratory participants performed at chance for all pairs, as expected. Note the flat performance across all pairs in the grey line in <xref ref-type="fig" rid="F1">Figure 1</xref>. This demonstrates that before training, participants were unable to discriminate between these unfamiliar speech sounds.</p>
<fig id="F1">
<label>Figure 1</label>
<caption>
<p>Average performance on Day 1 Pre-Test and Day 2 Post-Test for participants (n = 16) in the in-lab training condition. Error bars represent standard error of the mean.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-16-16542-g1.png"/>
</fig>
<p>After training, we observe a &#8220;peak&#8221; for stimulus pair 3&#8211;4 in the black line in <xref ref-type="fig" rid="F1">Figure 1</xref>, suggesting that in-laboratory participants have learned to differentiate between sounds that cross the trained category boundary, but not sounds within a single category. These results replicate decades of findings that after in-laboratory training, participants can learn to differentiate between two unfamiliar speech sounds.</p>
</sec>
<sec>
<title>3.2. Learning after remote training</title>
<p>When examining performance at pre-test for the remote participants, they perform at chance for all pairs, again as expected (see <xref ref-type="fig" rid="F2">Figure 2</xref>). Again, this demonstrates that participants were unable to discriminate between these speech sounds before training.</p>
<fig id="F2">
<label>Figure 2</label>
<caption>
<p>Average performance on Day 1 Pre-Test and Day 2 Post-Test for participants (n = 22) in the remote training condition. Error bars represent standard error of the mean.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-16-16542-g2.png"/>
</fig>
<p>At post-test, we again observe a &#8220;peak&#8221; at stimulus pair 3&#8211;4, showing that the remote participants learned to differentiate between sounds that cross the trained category boundary but not sounds within a single category (see <xref ref-type="fig" rid="F2">Figure 2</xref>).</p>
<p>These results show that participants can learn novel speech sounds in remote experimental set ups. Below, we statistically compare learning across the two modalities.</p>
</sec>
<sec>
<title>3.3. Comparison of learning</title>
<p>Examining <xref ref-type="fig" rid="F1">Figures 1</xref> and <xref ref-type="fig" rid="F2">2</xref>, it is clear that participants in the in-laboratory condition show a higher discrimination peak than participants in the remote condition. Therefore, we used the logistic mixed effects model described above to investigate the role of training type on performance by predicting proportion correct as a function of test phase, pair, training type, and their interactions. The model is summarized in <xref ref-type="table" rid="T1">Table 1</xref> below.</p>
<table-wrap id="T1">
<label>Table 1</label>
<caption>
<p>Model summary for logistic mixed effects model.</p>
</caption>
<table>
<thead>
<tr>
<td align="left" valign="top"><bold>Fixed Effect</bold></td>
<td align="left" valign="top"><bold>Estimate</bold></td>
<td align="left" valign="top"><bold>Standard Error</bold></td>
<td align="left" valign="top"><bold><bold><italic>z</italic></bold></bold></td>
<td align="left" valign="top"><bold><bold><italic>p</italic></bold></bold></td>
</tr>
</thead>
<tbody>
<tr>
<td align="left" valign="top">(Intercept)</td>
<td align="left" valign="top">0.234</td>
<td align="left" valign="top">0.078</td>
<td align="left" valign="top">3.025</td>
<td align="left" valign="top">0.002</td>
</tr>
<tr>
<td align="left" valign="top">Pair</td>
<td align="left" valign="top">0.368</td>
<td align="left" valign="top">0.092</td>
<td align="left" valign="top">4.003</td>
<td align="left" valign="top">&lt;.001</td>
</tr>
<tr>
<td align="left" valign="top">Training Type</td>
<td align="left" valign="top">&#8211;0.169</td>
<td align="left" valign="top">0.101</td>
<td align="left" valign="top">&#8211;1.679</td>
<td align="left" valign="top">0.093</td>
</tr>
<tr>
<td align="left" valign="top">Test Phase</td>
<td align="left" valign="top">&#8211;0.184</td>
<td align="left" valign="top">0.073</td>
<td align="left" valign="top">&#8211;2.518</td>
<td align="left" valign="top">0.012</td>
</tr>
<tr>
<td align="left" valign="top">Pair &#215; Training Type</td>
<td align="left" valign="top">&#8211;0.090</td>
<td align="left" valign="top">0.119</td>
<td align="left" valign="top">&#8211;0.761</td>
<td align="left" valign="top">0.448</td>
</tr>
<tr>
<td align="left" valign="top">Pair &#215; Test Phase</td>
<td align="left" valign="top">&#8211;0.155</td>
<td align="left" valign="top">0.129</td>
<td align="left" valign="top">&#8211;1.203</td>
<td align="left" valign="top">0.229</td>
</tr>
<tr>
<td align="left" valign="top">Training Type &#215; Test Phase</td>
<td align="left" valign="top">0.190</td>
<td align="left" valign="top">0.095</td>
<td align="left" valign="top">2.001</td>
<td align="left" valign="top">0.045</td>
</tr>
<tr>
<td align="left" valign="top">Pair &#215; Training Type &#215; Test Phase</td>
<td align="left" valign="top">&#8211;0.062</td>
<td align="left" valign="top">0.167</td>
<td align="left" valign="top">&#8211;0.373</td>
<td align="left" valign="top">0.70</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>First, we see that pair is significant, demonstrating that participants discriminate between tokens 3 and 4 more accurately than other pairs along the continuum, which suggests categorical perception. Next, we see that the comparison between in-laboratory training and remote training is also significant, showing that participants in the in-laboratory training do, indeed, perform better than participants in the remote training group. Further, we see that test phase (i.e., Day 1 Pre-Test vs. Day 2 Post-Test) is significant, demonstrating a change in participant performance from Day 1 Pre-Test to Day 2 Post-Test. The interactions for pair by training type and pair by test phase are not significant; however, the interaction between training type and test phase is. This captures the observation above that participants in the in-laboratory training group are more accurate than participants in the remote training group at post-test, though they do not differ at pre-test. The three-way interaction is not significant.<xref ref-type="fn" rid="n5">5</xref> It is possible that the lack of significant interactions is driven by individual differences among participants, and that investigating that variability will be informative. These statistical results suggest that participants do learn during training, and performance for in-laboratory participants is superior to remote participants. Below we explore in more detail why we may see these results.</p>
<sec>
<title>3.3.1. Learning for individual participants</title>
<p>Previous work has demonstrated that individual performance among in-laboratory participants on this type of learning task differs significantly (e.g., <xref ref-type="bibr" rid="B4">Baese-Berk, 2019</xref>; <xref ref-type="bibr" rid="B57">Perrachione et al., 2011</xref>). In this section, we investigate individual performance on the post-test.</p>
<p>Specifically, we ask what proportion of participants demonstrate improvement from Day 1 Pre-Test to Day 2 Post-Test. For the purposes of this exploration, we define improvement from Day 1 to Day 2 as a change of .1 or greater on proportion of correct trials for the 3&#8211;4 pair. For the in-laboratory participants, 8 of 16 participants demonstrate a clear improvement from Day 1 to Day 2 for the 3&#8211;4 pair (see <xref ref-type="fig" rid="F3">Figure 3</xref>). For the remote participants, only 9 of the 22 participants demonstrate a clear improvement (see <xref ref-type="fig" rid="F4">Figure 4</xref>). In both cases, other participants demonstrate myriad other patterns, but do not demonstrate the canonical peak seen in the group data. A Fischer&#8217;s exact test demonstrates that there is not a significant difference between the proportion of participants who learn (or demonstrate the canonical peak) across the two conditions (<italic>p</italic> = .397). This suggests that there is no evidence in this data that one type of training results in a higher fraction of individuals learning than the other.</p>
<fig id="F3">
<label>Figure 3</label>
<caption>
<p>Individual performance on the Day 1 Pre-Test and Day 2 Post-Test for participants in the in-laboratory training condition.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-16-16542-g3.png"/>
</fig>
<fig id="F4">
<label>Figure 4</label>
<caption>
<p>Individual performance on the Day 1 Pre-Test and Day 2 Post-Test for participants in the remote training condition.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-16-16542-g4.png"/>
</fig>
</sec>
<sec>
<title>3.3.2. Variability in performance</title>
<p>Next, we ask whether the variability in individual performance is different across the two groups. We again explored the difference in performance from Day 1 Pre-Test to Day 2 Post-test as a function of training condition. However, rather than asking a binary question of whether participants improve or not, we instead ask what patterns the group of participants show in terms of how much learning (or not) is demonstrated (see <xref ref-type="fig" rid="F5">Figure 5</xref>).</p>
<fig id="F5">
<label>Figure 5</label>
<caption>
<p>Difference between Day 1 Pre-Test and Day 2 Post-Test performance on the 3&#8211;4 pair for the in-laboratory and remote groups.</p>
</caption>
<graphic xmlns:xlink="http://www.w3.org/1999/xlink" xlink:href="labphon-16-16542-g5.png"/>
</fig>
<p>Our initial hypothesis regarding variability was that the remote group would be more variable because they represent a more heterogeneous group of individuals. That is, participants who completed the experiment in the lab were all enrolled as college students at the University of Oregon and likely shared many demographic features. However, participants in the remote group were recruited from across the United States and likely varied more in their backgrounds, including in their current educational status. Interestingly, however, our data does not appear to be consistent with the original hypothesis. Using Levene&#8217;s test for equality of variances, there is no difference between the two groups (<italic>F</italic> = 4.0355, <italic>p</italic> = .052). Examining <xref ref-type="fig" rid="F5">Figure 5</xref>, we see that, if anything, the in-laboratory group demonstrates more variable performance than the remote group, not the predicted reverse pattern. Indeed, all participants in the remote group have counterparts in the in-lab group who demonstrate similar performance to them. We discuss the implications of this finding in more detail in the Discussion section, below.</p>
</sec>
</sec>
<sec>
<title>3.4. Comparison of logistical issues during training</title>
<p>When evaluating the two conditions for training, it is important to evaluate not only learning performance but also the logistics of conducting such experiments. One key consideration is the time it takes to recruit participants. In the case of the in-laboratory study, we recruited 18 participants in a single 10-week quarter. Sixteen of these participants completed the study. While myriad factors impact our ability to recruit participants, this recruitment pattern was quite typical for our lab for a two-day training study. It is often quite easy to recruit participants for shorter studies (i.e., single day studies), but for multiple day studies recruitment is typically much slower.</p>
<p>In contrast, participants in the remote condition were recruited on a single day. In fact, it only took three hours to recruit 30 participants, 22 of whom completed the study. Because the experiment took two days to complete, the data collection also took two days, but this time period was much shorter than the 10-week period described above. Indeed, if we had aimed for a much larger sample size, we could have recruited even more participants in the aforementioned three-hour time period.</p>
<p>It is also important to compare attrition rates across the two studies. In the in-laboratory condition, 2 of 18 participants did not return for the second day (around 11% of participants). In the remote condition, 8 of 30 participants did not complete the second day of training (27% of participants). While the attrition rate is higher for the remote condition, it is not remarkably higher, as attrition rates in our lab are often between 15 and 25% for multi-day training studies. Therefore, we do not believe that this is a significant disadvantage for the logistics of running such a study in a remote setting.</p>
<p>There were not significant differences between the groups in how long it took them to complete the task; all participants in both groups took around one hour each day to complete the task. Therefore, we do not believe that overall group differences in performance can be attributable to time-on-task.</p>
<p>One final question is one of demographics. In our data, our in-laboratory participants all identified as White (n = 16), and the majority of participants identified as female (n = 10; male n = 5; non-binary n = 1). This is not surprising given the demographics of the university and specifically of the Linguistics and Psychology Human Subject Pool we used to recruit participants.<xref ref-type="fn" rid="n6">6</xref> The self-reported demographics of our remote participants were more ethnically diverse (White n = 18; Black n = 2; Asian n = 1; multiple racial ethnicities n = 1). The gender demographics were skewed more heavily toward male participants than our in-laboratory participants (male n = 13; female n = 7; prefer not to say = 2). Though we did not ask in our demographic questionnaire, it is also possible that our participant pools differed in other demographic characteristics including socioeconomic and educational status. The differences reported here between our in-lab and remote participant demographics are not substantial. However, it is important to note that it is possible to actively recruit for more diverse populations in remote studies, which is more challenging for in-laboratory studies, especially in locations where the general population and, specifically, the university population of question, are more homogeneous.</p>
</sec>
</sec>
<sec>
<title>4. Discussion</title>
<p>In this study, we examined learning of a novel phonological contrast after training conducted either in-laboratory or remotely. Our results demonstrate that, on a group level, both sets of participants demonstrated significant improvement from pre-test to post-test after training. However, participants in the remote condition demonstrated less learning than participants in the in-laboratory condition. Interestingly, participants in the remote condition were not more variable in their performance than participants in the in-laboratory condition, in spite of initial predictions to the contrary.</p>
<p>Our results demonstrate that training paradigms for novel speech sounds can be conducted remotely. As a group, our participants in both conditions improve from pre- to post-test. However, learning in the remote condition is less than that of the in-lab condition, which could result in some caution for researchers hoping to conduct such studies, especially given the wide individual variability among participants in both groups. However, in addition to some participants completing a remote training paradigm and others doing so in-laboratory, there are additional differences between the two populations here. This said, if comparison is within individuals participating in the same conditions (e.g., remote participants compared only to remote participants, rather than comparing in-lab participants to remote participants), it is likely that these differences do not preclude conducting experiments remotely, especially when the benefits of conducting experiments remotely are quite high (e.g., situations where data must be collected quickly or from populations not easily attainable for in-laboratory studies).</p>
<p>An additional factor to consider is that the population in the in-laboratory study is quite homogeneous. All participants were current students enrolled at the same university and who lived in the same geographic area. Therefore, education level was controlled and various other factors including race, socioeconomic status and personal background were unlikely to vary substantially given the demographics of both the region and institution. On the other hand, our remote participants were from across the United States, or at least were individuals using an IP address located within the United States. While they were roughly matched with the in-laboratory participants for age, all other factors were likely more variable for the remote participants. However, differences in demographics cannot fully account for differences in performance. For example, other studies that have matched in-laboratory and remote participant groups for a series of speech perception tasks have found that participants from the same population perform slightly differently in the two types of settings (e.g., <xref ref-type="bibr" rid="B22">Cooke &amp; Garc&#237;a Lecumberri, 2021</xref>).</p>
<p>As researchers who often crave control of our experimental population such that we do not introduce unnecessary variability in our data, this diversity may be a bit disconcerting. However, it is important to note that increasing the demographic variability in our population did not result in increased performance variability in the remote condition. Further, if the goal of psycholinguistic and laboratory phonology work is to capture generalizations about behavior or about language, we should question what it means when our results only hold for some subset of a population. That is, if our sample is not truly representative, what might this mean for our ability to make generalizations? Many recent papers have argued for more inclusive approaches to psychological and linguistic research and teaching (<xref ref-type="bibr" rid="B6">Baese-Berk &amp; Reed, 2023</xref>; <xref ref-type="bibr" rid="B34">Higby et al., 2023</xref>; <xref ref-type="bibr" rid="B42">Kirk, 2023</xref>; <xref ref-type="bibr" rid="B45">Kutlu &amp; Hayes-Harb, 2023</xref>; <xref ref-type="bibr" rid="B52">McMurray et al., 2023</xref>; <xref ref-type="bibr" rid="B59">Rad et al., 2018</xref>; <xref ref-type="bibr" rid="B69">Tripp &amp; Munson, 2023</xref>). That said, increasing our samples beyond those which are most convenient (e.g., college students at our institutions) can feel daunting to researchers who are used to conducting in-laboratory studies with the convenient samples of individuals within their community. Using remote data collection protocols allows for relatively easy access to diverse populations. Indeed, in addition to collecting data from the general American public, remote data collection allows for research with a variety of specific populations. For example, collecting data from Spanish Heritage speakers in person was a very arduous process in our lab; however, when switching to remote data collection, we collected data from many such speakers, and these individuals had a substantially more heterogeneous background than those available to visit our laboratory in person (<xref ref-type="bibr" rid="B65">Staggs et al., 2022</xref>).</p>
<p>As discussed above, the logistics of conducting research remotely are often much easier than doing so in person. It can be challenging to schedule multi-day studies for participants at times that are also convenient or available for researchers. This is important because in addition to alleviating burdens on researchers in general, remote data collection may also allow for researchers who themselves are marginalized in a variety of ways to conduct research more easily. That is, individuals who are not at large institutions with human subject pools may find remote data collection allows them to complete studies relatively quickly that might otherwise take years to complete.</p>
<p>Even given all of the benefits of remote data collection described above, one could examine our data and still express concern that participants in the remote training condition performed less well than in-laboratory participants. That is, could this be a sign of less high-quality data in the remote group? For example, it is possible that participants in the remote condition are more distracted than participants in the in-laboratory condition. That is, participants in a laboratory setting tend to have a more controlled environment with fewer variables that may take a participants&#8217; attention away from the task at hand (see, e.g., <xref ref-type="bibr" rid="B2">Aivaz &amp; Teodorescu, 2022</xref>). For example, smart phones have been shown to have a detrimental, distracting effect on learning in classrooms (<xref ref-type="bibr" rid="B25">Dontre, 2021</xref>), and the availability of such devices during remote learning could be responsible for the decrement in learning seen here. Alternatively, it could be the case that, in addition to demographic differences, remote participants are less accustomed to being in an educational or learning-oriented setting than our typical in-laboratory participants (e.g., <xref ref-type="bibr" rid="B9">Belot et al., 2015</xref>; <xref ref-type="bibr" rid="B35">Hooghe et al., 2010</xref>), and this lack of (recent) experience with educational settings and testing impacts performance.</p>
<p>We suggest that rather than viewing the decreased performance for remote participants compared to in-laboratory participants as a caution or a warning sign, we should view this as data&#8212;evidence that perhaps there are opportunities for interesting questions about how different groups of learners may perform differently, as well as what types of training might result in the most robust learning for different types of learners.</p>
<p>While the data presented here cannot differentiate between the effects of training modality and population differences, they do provide an opportunity to develop new research questions around how learners best acquire novel speech sounds. However, these questions are not just methodological. Theoretical questions can also be addressed by comparing both training conditions and participant populations. For example, significant previous work has suggested that individual variability in speech sound learning is not just unexplainable noise, but in fact correlates with various cognitive properties of the participant (e.g., working memory; <xref ref-type="bibr" rid="B50">McHaney et al., 2021</xref>; <xref ref-type="bibr" rid="B60">Roark et al., 2022</xref>; <xref ref-type="bibr" rid="B57">Perrachione et al., 2011</xref>). In some recent work, we have questioned whether all participants are, indeed, learning the same things during speech sound training tasks (<xref ref-type="bibr" rid="B5">Baese-Berk et al., 2022</xref>). That is, there is known to be substantial individual variation in performance both before, during, and after training on differentiating or categorizing novel speech sounds. Learners differ not only in performance but also strategies used (e.g., <xref ref-type="bibr" rid="B19">Chandrasekaran et al., 2014</xref>), weighting a variety of cues (e.g., <xref ref-type="bibr" rid="B62">Schertz et al., 2015</xref>), and various cognitive properties which may impact learning (e.g., <xref ref-type="bibr" rid="B32">Heffner &amp; Myers, 2021</xref>). This leads to a question of whether participants in all cases are acquiring novel categories, and if not, what participants might be learning during training.</p>
<p>Indeed, this issue brings forth an even broader question about the nature of the assumptions we make in conducting studies like these. That is, a growing body of work questions the notion of categorical perception for speech (e.g., <xref ref-type="bibr" rid="B51">McMurray, 2022</xref>). Perhaps the variation we see in participant performance here could be informative about precisely what learners are acquiring and what our observations of how people learn can tell us about categorical structure, including interfaces of phonetics and phonology, broadly speaking. It is also possible that the use of a single pair (i.e., 3&#8211;4 in this study) is problematic because for some participants another pair crosses the category boundary instead. Indeed, if this is the case, one would expect less learning on a group level because good discrimination on one pair for one participant might be &#8220;canceled out&#8221; by poor performance by another participant. While this doesn&#8217;t appear to be the case in our data, it is possible that different participants have different category boundaries. This possibility is not often addressed in the literature because many studies assume (and have demonstrated) a natural psychophysical boundary for some contrasts (e.g., <xref ref-type="bibr" rid="B28">Elangovan &amp; Stuart, 2008</xref>). While the effect of task on categorical perception has been well-studied (e.g., <xref ref-type="bibr" rid="B38">Kapnoula &amp; McMurray, 2021</xref>; <xref ref-type="bibr" rid="B58">Pisoni &amp; Tash, 1974</xref>; <xref ref-type="bibr" rid="B64">Schouten et al., 2003</xref>), the issue of addressing potential variance in individual performance as a function of different category boundaries, especially in cases of learning novel contrasts, is less understood. The consensus in the field seems to be shifting toward a need to better understand <italic>what</italic> precisely is being learned, and it is possible that using remote experiments will help address these issues.</p>
<sec>
<title>4.1. Recommendations for conducting remote experiments</title>
<p>Taken together, the results of this study demonstrate the positive and negative aspects of conducting research remotely. While we are not the first to consider the trade-offs of these two means of data collection (see, e.g., <xref ref-type="bibr" rid="B27">Eerola et al., 2021</xref>), few studies have directly examined the pros and cons of conducting multi-day studies in these modalities and none has focused specifically on learning novel speech sounds. Below, we put forth a few recommendations for researchers deciding whether to conduct multi-day training studies remotely vs. in person.</p>
<p>First, researchers should consider logistical issues. If it is difficult to recruit participants to multi-day in-laboratory studies, or if they are hoping to target a specific population of participants (or a more diverse population than is available in their local area), remote data collection can provide a promising alternative to in-laboratory data collection.</p>
<p>Next, researchers should consider a number of methodological and analytical choices before beginning to conduct their research. In the present study, attrition was higher in the remote condition than the in-laboratory condition. How will the researcher handle attrition in the sample? Will they exclude participants who do not complete the entire experiment? Will they replace those who choose not to complete the study to ensure sufficient power? On the issue of power, the experimenter should consider whether a larger sample size is necessary to generate sufficient statistical power, as effect sizes may be smaller for remote experiments than for in-laboratory experiments.</p>
<p>The issue of potential distraction is also key to consider. How will a researcher determine whether a participant was too distracted? Or, more basically, how much does attention to task matter for the question being posed in a specific study? One could imagine a variety of controls for attention, including attention checks throughout the study, timed responses requiring participants to engage carefully on a given trial, or even use of remote eye-tracking tasks to note whether participants are looking at the screen during a given trial. The closer each task moves toward surveillance (e.g., a live experimenter being present via videoconference, for example), the more taxing this becomes for a laboratory and the more invasive the task becomes for participants (<xref ref-type="bibr" rid="B17">Castelli &amp; Sarvary, 2021</xref>). This suggests a trade-off between controlled experimental settings and ease of remote data collection.</p>
<p>Further, the researchers should think carefully about how they collect data about their participants and whether verification of this data is mandatory in order to properly interpret their results. For example, it may be more difficult to verify some aspects of a participant&#8217;s language background online than it is in person. If having a clear understanding of language background is mandatory for a specific experiment, researchers should consider whether additional measures may be required to ensure participants have a specific language background. However, it is also important to consider whether such information is truly necessary for a particular study, given that &#8220;nativeness,&#8221; for example, is not a simple construct and may not impact study results as we expect (<xref ref-type="bibr" rid="B20">Cheng et al., 2021</xref>; <xref ref-type="bibr" rid="B67">Strand et al., 2024</xref>).</p>
<p>Given all of this, researchers should consider whether conducting data collection in a remote setting might be expected to impact their results in a meaningful way. That is, does shifting to remote data collection so drastically change the conditions for experimentation that it becomes a research question in and of itself?</p>
<p>Finally, and most crucially, researchers should be clear about their methodological and analytical decisions when reporting their results to ensure that results across studies are comparable and that results in any given paper are replicable.</p>
</sec>
</sec>
<sec>
<title>5. Conclusion</title>
<p>In the present study, we demonstrate that participants in both in-laboratory and remote experimental settings can learn a novel speech sound distinction. However, differences in performance between the two groups suggest that remote and in-laboratory conditions are not identical and warrant significant consideration before replacing one with the other. While we believe that conducting laboratory phonology experiments remotely could improve representation of diversity in the populations we work with, it is not something that can be undertaken without significant consideration for how the change to remote settings may impact the results and interpretation of a given study.</p>
</sec>
</body>
<back>
<fn-group>
<fn id="n1"><p>Note that remote data collection may have similar challenges in terms of representation, given that individuals who seek out online studies may also not be representative of a broader population.</p></fn>
<fn id="n2"><p>Interestingly some recent work suggests that use of headphones may not drastically impact some perception data (<xref ref-type="bibr" rid="B61">Sanker, 2023</xref>).</p></fn>
<fn id="n3"><p>The key correspondence to response was presented on the screen for each trial and has been previously used in other studies (e.g., <xref ref-type="bibr" rid="B7">Baese-Berk &amp; Samuel, 2016</xref>).</p></fn>
<fn id="n4"><p>It should be noted that there has been substantial debate recently about the nature of categorical perception in general (e.g., <xref ref-type="bibr" rid="B51">McMurray, 2022</xref>) and in second language speech sound learning more specifically (<xref ref-type="bibr" rid="B5">Baese-Berk, Chandrasekaran, &amp; Roark, 2022</xref>), including some work suggesting that learners are not actually acquiring categories. Further, the use of a single pair as the base pair has a long history; however, it is not without its own problems. We will return to these issues in the discussion section.</p></fn>
<fn id="n5"><p>Please note that the conclusions one can draw from these simple effects are limited by the fact that simple effects apply at the default level of other variables. That is, the improvement from pre- to post-test holds for the in-laboratory group at pair 3&#8211;4 (i.e., the default levels for those two variables).</p></fn>
<fn id="n6"><p>Note that although the human subject pool used here included students from both linguistics and psychology classes, the students enrolled in this specific study were all currently enrolled in psychology classes and did not report experience in linguistic classes.</p></fn>
</fn-group>
<sec>
<title>Acknowledgements</title>
<p>This work is supported by NSF Grants BCS-2117665 and IIS-2024926. We would also like to thank Zachary Jaggers for his help in programming these experiments and Kurtis Foster for his assistance in data collection.</p>
</sec>
<sec>
<title>Competing interests</title>
<p>The authors have no competing interests to declare.</p>
</sec>
<ref-list>
<ref id="B1"><mixed-citation publication-type="journal"><string-name><surname>Aguinis</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Villamor</surname>, <given-names>I.</given-names></string-name>, &amp; <string-name><surname>Ramani</surname>, <given-names>R.</given-names></string-name> (<year>2021</year>). <article-title>Mturk research: Review and recommendations</article-title>. <source>Journal of Management</source>, <volume>47</volume>(<issue>4</issue>), <fpage>823</fpage>&#8211;<lpage>837</lpage>.</mixed-citation></ref>
<ref id="B2"><mixed-citation publication-type="journal"><string-name><surname>Aivaz</surname>, <given-names>K. A.</given-names></string-name>, &amp; <string-name><surname>Teodorescu</surname>, <given-names>D.</given-names></string-name> (<year>2022</year>). <article-title>College students&#8217; distractions from learning caused by multitasking in online vs. face-to-face classes: A case study at a public university in Romania</article-title>. <source>International Journal of Environmental Research and Public Health</source>, <volume>19</volume>(<issue>18</issue>), <elocation-id>11188</elocation-id>.</mixed-citation></ref>
<ref id="B3"><mixed-citation publication-type="journal"><string-name><surname>Andringa</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Godfroid</surname>, <given-names>A.</given-names></string-name> (<year>2020</year>). <article-title>Sampling bias and the problem of generalizability in applied linguistics</article-title>. <source>Annual Review of Applied Linguistics</source>, <volume>40</volume>, <fpage>134</fpage>&#8211;<lpage>142</lpage>. <pub-id pub-id-type="doi">10.1017/S0267190520000033</pub-id></mixed-citation></ref>
<ref id="B4"><mixed-citation publication-type="journal"><string-name><surname>Baese-Berk</surname>, <given-names>M. M.</given-names></string-name> (<year>2019</year>). <article-title>Interactions between speech perception and production during learning of novel phonemic categories</article-title>. <source>Attention, Perception, &amp; Psychophysics</source>, <volume>81</volume>(<issue>4</issue>), <fpage>981</fpage>&#8211;<lpage>1005</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-019-01725-4</pub-id></mixed-citation></ref>
<ref id="B5"><mixed-citation publication-type="journal"><string-name><surname>Baese-Berk</surname>, <given-names>M. M.</given-names></string-name>, <string-name><surname>Chandrasekaran</surname>, <given-names>B.</given-names></string-name>, &amp; <string-name><surname>Roark</surname>, <given-names>C. L.</given-names></string-name> (<year>2022</year>). <article-title>The nature of non-native speech sound representations</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>152</volume>(<issue>5</issue>), <fpage>3025</fpage>&#8211;<lpage>3034</lpage>. <pub-id pub-id-type="doi">10.1121/10.0015230</pub-id></mixed-citation></ref>
<ref id="B6"><mixed-citation publication-type="journal"><string-name><surname>Baese-Berk</surname>, <given-names>M. M.</given-names></string-name>, &amp; <string-name><surname>Reed</surname>, <given-names>P. E.</given-names></string-name> (<year>2023</year>). <article-title>Addressing diversity in speech science courses</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>154</volume>(<issue>2</issue>), <fpage>918</fpage>&#8211;<lpage>925</lpage>.</mixed-citation></ref>
<ref id="B7"><mixed-citation publication-type="journal"><string-name><surname>Baese-Berk</surname>, <given-names>M. M.</given-names></string-name>, &amp; <string-name><surname>Samuel</surname>, <given-names>A. G.</given-names></string-name> (<year>2016</year>). <article-title>Listeners beware: Speech production may be bad for learning speech sounds</article-title>. <source>Journal of Memory and Language</source>, <volume>89</volume>, <fpage>23</fpage>&#8211;<lpage>36</lpage>.</mixed-citation></ref>
<ref id="B8"><mixed-citation publication-type="journal"><string-name><surname>Baese-Berk</surname>, <given-names>M. M.</given-names></string-name>, &amp; <string-name><surname>Samuel</surname>, <given-names>A. G.</given-names></string-name> (<year>2022</year>). <article-title>Just give it time: Differential effects of disruption and delay on perceptual learning</article-title>. <source>Attention, Perception, &amp; Psychophysics</source>, <volume>84</volume>(<issue>3</issue>), <fpage>960</fpage>&#8211;<lpage>980</lpage>. <pub-id pub-id-type="doi">10.3758/s13414-022-02463-w</pub-id></mixed-citation></ref>
<ref id="B9"><mixed-citation publication-type="journal"><string-name><surname>Belot</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Duch</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Miller</surname>, <given-names>L.</given-names></string-name> (<year>2015</year>). <article-title>A comprehensive comparison of students and non-students in classic experimental games</article-title>. <source>Journal of Economic Behavior &amp; Organization</source>, <volume>113</volume>, <fpage>26</fpage>&#8211;<lpage>33</lpage>.</mixed-citation></ref>
<ref id="B10"><mixed-citation publication-type="journal"><string-name><surname>Bent</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Lind-Combs</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Holt</surname>, <given-names>R. F.</given-names></string-name>, &amp; <string-name><surname>Clopper</surname>, <given-names>C.</given-names></string-name> (<year>2024</year>). <article-title>Perception of regional and nonnative accents: A comparison of museum laboratory and online data collection</article-title>. <source>Linguistics Vanguard</source>, <volume>9</volume>(<issue>s4</issue>), <fpage>361</fpage>&#8211;<lpage>373</lpage>. <pub-id pub-id-type="doi">10.1515/lingvan-2021-0157</pub-id></mixed-citation></ref>
<ref id="B11"><mixed-citation publication-type="book"><string-name><surname>Best</surname>, <given-names>C. T.</given-names></string-name>, &amp; <string-name><surname>Tyler</surname>, <given-names>M. D.</given-names></string-name> (<year>2007</year>). <chapter-title>Nonnative and second-language speech perception: Commonalities and complementaries</chapter-title>. In <string-name><given-names>O. S.</given-names> <surname>Bohn</surname></string-name> (Ed.), <source>Language experience in second language speech learning: In honor of James Emil Flege</source> (pp. <fpage>13</fpage>&#8211;<lpage>34</lpage>). <publisher-name>John Benjamins</publisher-name>.</mixed-citation></ref>
<ref id="B12"><mixed-citation publication-type="journal"><string-name><surname>Bradlow</surname>, <given-names>A. R.</given-names></string-name>, <string-name><surname>Akahane-Yamada</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Pisoni</surname>, <given-names>D. B.</given-names></string-name>, &amp; <string-name><surname>Tohkura</surname>, <given-names>Y.</given-names></string-name> (<year>1999</year>). <article-title>Training Japanese listeners to identify English /r/and /l/: Long-term retention of learning in perception and production</article-title>. <source>Perception Psychophysics</source>, <volume>61</volume>(<issue>5</issue>), <fpage>977</fpage>&#8211;<lpage>985</lpage>.</mixed-citation></ref>
<ref id="B13"><mixed-citation publication-type="journal"><string-name><surname>Bradlow</surname>, <given-names>A. R.</given-names></string-name>, <string-name><surname>Pisoni</surname>, <given-names>D. B.</given-names></string-name>, <string-name><surname>Akahane-Yamada</surname>, <given-names>R.</given-names></string-name>, &amp; <string-name><surname>Tohkura</surname>, <given-names>Y.</given-names></string-name> (<year>1997</year>). <article-title>Training Japanese listeners to identify English/r/and/l: IV. Some effects of perceptual learning on speech production</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>101</volume>(<issue>4</issue>), <fpage>2299</fpage>&#8211;<lpage>2310</lpage>. <pub-id pub-id-type="doi">10.1121/1.418276</pub-id></mixed-citation></ref>
<ref id="B14"><mixed-citation publication-type="journal"><string-name><surname>Brekelmans</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Lavan</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Saito</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Clayards</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Wonnacott</surname>, <given-names>E.</given-names></string-name> (<year>2022</year>). <article-title>Does high variability training improve the learning of non-native phoneme contrasts over low variability training? A replication</article-title>. <source>Journal of Memory and Language</source>, <volume>126</volume>, <elocation-id>104352</elocation-id>.</mixed-citation></ref>
<ref id="B15"><mixed-citation publication-type="journal"><string-name><surname>Bro&#347;</surname>, <given-names>K.</given-names></string-name> (<year>2025</year>). <article-title>Remote data collection in the study of ongoing sound change in Spanish &#8211; a comparative analysis</article-title>. <source>Laboratory Phonology</source>, <volume>16</volume>(<issue>1</issue>). <pub-id pub-id-type="doi">10.16995/labphon.10557</pub-id></mixed-citation></ref>
<ref id="B16"><mixed-citation publication-type="journal"><string-name><surname>Buhrmester</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Kwang</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Gosling</surname>, <given-names>S. D.</given-names></string-name> (<year>2011</year>). <article-title>Amazon&#8217;s Mechanical Turk: A new source of inexpensive, yet high-quality, data?</article-title> <source>Perspectives on Psychological Science</source>, <volume>6</volume>(<issue>1</issue>), <fpage>3</fpage>&#8211;<lpage>5</lpage>. <pub-id pub-id-type="doi">10.1177/1745691610393980</pub-id></mixed-citation></ref>
<ref id="B17"><mixed-citation publication-type="journal"><string-name><surname>Castelli</surname>, <given-names>F. R.</given-names></string-name>, &amp; <string-name><surname>Sarvary</surname>, <given-names>M. A.</given-names></string-name> (<year>2021</year>). <article-title>Why students do not turn on their video cameras during online classes and an equitable and inclusive plan to encourage them to do so</article-title>. <source>Ecology and Evolution</source>, <volume>11</volume>(<issue>8</issue>), <fpage>3565</fpage>&#8211;<lpage>3576</lpage>.</mixed-citation></ref>
<ref id="B18"><mixed-citation publication-type="journal"><string-name><surname>Chandler</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Paolacci</surname>, <given-names>G.</given-names></string-name> (<year>2017</year>). <article-title>Lie for a dime: When most prescreening responses are honest but most study participants are imposters</article-title>. <source>Social Psychological and Personality Science</source>, <volume>8</volume>(<issue>5</issue>), <fpage>500</fpage>&#8211;<lpage>508</lpage>.</mixed-citation></ref>
<ref id="B19"><mixed-citation publication-type="journal"><string-name><surname>Chandrasekaran</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Yi</surname>, <given-names>H.-G.</given-names></string-name>, &amp; <string-name><surname>Maddox</surname>, <given-names>W. T.</given-names></string-name> (<year>2014</year>). <article-title>Dual-learning systems during speech category learning</article-title>. <source>Psychonomic Bulletin &amp; Review</source>, <volume>21</volume>(<issue>2</issue>), <fpage>488</fpage>&#8211;<lpage>495</lpage>. <pub-id pub-id-type="doi">10.3758/s13423-013-0501-5</pub-id></mixed-citation></ref>
<ref id="B20"><mixed-citation publication-type="journal"><string-name><surname>Cheng</surname>, <given-names>L. S.</given-names></string-name>, <string-name><surname>Burgess</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Vernooij</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Sol&#237;s-Barroso</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>McDermott</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Namboodiripad</surname>, <given-names>S.</given-names></string-name> (<year>2021</year>). <article-title>The problematic concept of native speaker in psycholinguistics: Replacing vague and harmful terminology with inclusive and accurate measures</article-title>. <source>Frontiers in Psychology</source>, <volume>12</volume>, <elocation-id>715843</elocation-id>.</mixed-citation></ref>
<ref id="B21"><mixed-citation publication-type="journal"><string-name><surname>Chodroff</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Wilson</surname>, <given-names>C.</given-names></string-name> (<year>2014</year>). <article-title>Burst spectrum as a cue for the stop voicing contrast in American English</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>136</volume>(<issue>5</issue>), <fpage>2762</fpage>&#8211;<lpage>2772</lpage>. <pub-id pub-id-type="doi">10.1121/1.4896470</pub-id></mixed-citation></ref>
<ref id="B22"><mixed-citation publication-type="journal"><string-name><surname>Cooke</surname>, <given-names>M.</given-names></string-name>, &amp; <string-name><surname>Garc&#237;a Lecumberri</surname>, <given-names>M. L.</given-names></string-name> (<year>2021</year>). <article-title>How reliable are online speech intelligibility studies with known listener cohorts?</article-title> <source>Journal of the Acoustical Society of America</source>, <volume>150</volume>(<issue>2</issue>), <fpage>1390</fpage>&#8211;<lpage>1401</lpage>.</mixed-citation></ref>
<ref id="B23"><mixed-citation publication-type="journal"><string-name><surname>Crump</surname>, <given-names>M. J.</given-names></string-name>, <string-name><surname>McDonnell</surname>, <given-names>J. V.</given-names></string-name>, &amp; <string-name><surname>Gureckis</surname>, <given-names>T. M.</given-names></string-name> (<year>2013</year>). <article-title>Evaluating Amazon&#8217;s Mechanical Turk as a tool for experimental behavioral research</article-title>. <source>PLOS One</source>, <volume>8</volume>(<issue>3</issue>), <elocation-id>e57410</elocation-id>. <pub-id pub-id-type="doi">10.1371/journal.pone.0057410</pub-id></mixed-citation></ref>
<ref id="B24"><mixed-citation publication-type="journal"><string-name><surname>Darcy</surname>, <given-names>I.</given-names></string-name>, <string-name><surname>Park</surname>, <given-names>H.</given-names></string-name>, &amp; <string-name><surname>Yang</surname>, <given-names>C.-L.</given-names></string-name> (<year>2015</year>). <article-title>Individual differences in L2 acquisition of English phonology: The relation between cognitive abilities and phonological processing</article-title>. <source>Learning and Individual Differences</source>, <volume>40</volume>, <fpage>63</fpage>&#8211;<lpage>72</lpage>. <pub-id pub-id-type="doi">10.1016/j.lindif.2015.04.005</pub-id></mixed-citation></ref>
<ref id="B25"><mixed-citation publication-type="journal"><string-name><surname>Dontre</surname>, <given-names>A. J.</given-names></string-name> (<year>2021</year>). <article-title>The influence of technology on academic distraction: A review</article-title>. <source>Human Behavior and Emerging Technologies</source>, <volume>3</volume>(<issue>3</issue>), <fpage>379</fpage>&#8211;<lpage>390</lpage>.</mixed-citation></ref>
<ref id="B26"><mixed-citation publication-type="journal"><string-name><surname>Earle</surname>, <given-names>F. S.</given-names></string-name>, &amp; <string-name><surname>Myers</surname>, <given-names>E. B.</given-names></string-name> (<year>2015</year>). <article-title>Sleep and native language interference affect non-native speech sound learning</article-title>. <source>Journal of Experimental Psychology: Human Perception and Performance</source>, <volume>41</volume>(<issue>6</issue>), <fpage>1680</fpage>&#8211;<lpage>1695</lpage>. <pub-id pub-id-type="doi">10.1037/xhp0000113</pub-id></mixed-citation></ref>
<ref id="B27"><mixed-citation publication-type="journal"><string-name><surname>Eerola</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Armitage</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Lavan</surname>, <given-names>N.</given-names></string-name>, &amp; <string-name><surname>Knight</surname>, <given-names>S.</given-names></string-name> (<year>2021</year>). <article-title>Online data collection in auditory perception and cognition research: Recruitment, testing, data quality and ethical considerations</article-title>. <source>Auditory Perception &amp; Cognition</source>, <volume>4</volume>(<issue>3&#8211;4</issue>), <fpage>251</fpage>&#8211;<lpage>280</lpage>. <pub-id pub-id-type="doi">10.1080/25742442.2021.2007718</pub-id></mixed-citation></ref>
<ref id="B28"><mixed-citation publication-type="journal"><string-name><surname>Elangovan</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Stuart</surname>, <given-names>A.</given-names></string-name> (<year>2008</year>). <article-title>Natural boundaries in gap detection are related to categorical perception of stop consonants</article-title>. <source>Ear and Hearing</source>, <volume>29</volume>(<issue>5</issue>), <fpage>761</fpage>&#8211;<lpage>774</lpage>.</mixed-citation></ref>
<ref id="B29"><mixed-citation publication-type="journal"><string-name><surname>Enochson</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Culbertson</surname>, <given-names>J.</given-names></string-name> (<year>2015</year>). <article-title>Collecting psycholinguistic response time data using Amazon Mechanical Turk</article-title>. <source>PLOS One</source>, <volume>10</volume>(<issue>3</issue>), <elocation-id>e0116946</elocation-id>. <pub-id pub-id-type="doi">10.1371/journal.pone.0116946</pub-id></mixed-citation></ref>
<ref id="B30"><mixed-citation publication-type="book"><string-name><surname>Flege</surname>, <given-names>J. E.</given-names></string-name>, &amp; <string-name><surname>Bohn</surname>, <given-names>O.-S.</given-names></string-name> (<year>2021</year>). <chapter-title>The revised speech learning model (SLM-r)</chapter-title>. In <string-name><given-names>R.</given-names> <surname>Wayland</surname></string-name> (Ed.), <source>Second language speech learning: Theoretical and empirical progress</source> (pp. <fpage>3</fpage>&#8211;<lpage>83</lpage>). <publisher-name>Cambridge University Press</publisher-name>.</mixed-citation></ref>
<ref id="B31"><mixed-citation publication-type="journal"><string-name><surname>Grieve</surname>, <given-names>J.</given-names></string-name> (<year>2021</year>). <article-title>Observation, experimentation, and replication in linguistics</article-title>. <source>Linguistics</source>, <volume>59</volume>(<issue>5</issue>), <fpage>1343</fpage>&#8211;<lpage>1356</lpage>.</mixed-citation></ref>
<ref id="B32"><mixed-citation publication-type="journal"><string-name><surname>Heffner</surname>, <given-names>C. C.</given-names></string-name>, &amp; <string-name><surname>Myers</surname>, <given-names>E. B.</given-names></string-name> (<year>2021</year>). <article-title>Individual differences in phonetic plasticity across native and nonnative contexts</article-title>. <source>Journal of Speech, Language, and Hearing Research</source>, <volume>64</volume>(<issue>10</issue>), <fpage>3720</fpage>&#8211;<lpage>3733</lpage>. <pub-id pub-id-type="doi">10.1044/2021_JSLHR-21-00004</pub-id></mixed-citation></ref>
<ref id="B33"><mixed-citation publication-type="journal"><string-name><surname>Henrich</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Heine</surname>, <given-names>S. J.</given-names></string-name>, &amp; <string-name><surname>Norenzayan</surname>, <given-names>A.</given-names></string-name> (<year>2010</year>). <article-title>The weirdest people in the world?</article-title> <source>Behavioral and Brain Sciences</source>, <volume>33</volume>(<issue>2&#8211;3</issue>), <fpage>61</fpage>&#8211;<lpage>83</lpage>. <pub-id pub-id-type="doi">10.1017/S0140525X0999152X</pub-id></mixed-citation></ref>
<ref id="B34"><mixed-citation publication-type="journal"><string-name><surname>Higby</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>G&#225;mez</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Mendoza</surname>, <given-names>C. H.</given-names></string-name> (<year>2023</year>). <article-title>Challenging deficit frameworks in research on heritage language bilingualism</article-title>. <source>Applied Psycholinguistics</source>, <volume>44</volume>(<issue>4</issue>), <fpage>417</fpage>&#8211;<lpage>430</lpage>. <pub-id pub-id-type="doi">10.1017/S0142716423000048</pub-id>.</mixed-citation></ref>
<ref id="B35"><mixed-citation publication-type="journal"><string-name><surname>Hooghe</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Stolle</surname>, <given-names>D.</given-names></string-name>, <string-name><surname>Mah&#233;o</surname>, <given-names>V. A.</given-names></string-name>, &amp; <string-name><surname>Vissers</surname>, <given-names>S.</given-names></string-name> (<year>2010</year>). <article-title>Why can&#8217;t a student be more like an average person?: Sampling and attrition effects in social science field and laboratory experiments</article-title>. <source>The Annals of the American Academy of Political and Social Science</source>, <volume>628</volume>(<issue>1</issue>), <fpage>85</fpage>&#8211;<lpage>96</lpage>.</mixed-citation></ref>
<ref id="B36"><mixed-citation publication-type="journal"><string-name><surname>Houde</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>Jordan</surname>, <given-names>M.</given-names></string-name> (<year>2002</year>). <article-title>Sensorimotor adaptation of speech I: Compensation and adaptation</article-title>. <source>Journal of Speech, Language, and Hearing Research</source>, <volume>45</volume>, <fpage>295</fpage>&#8211;<lpage>310</lpage>.</mixed-citation></ref>
<ref id="B37"><mixed-citation publication-type="journal"><string-name><surname>Iverson</surname>, <given-names>P.</given-names></string-name>, <string-name><surname>Hazan</surname>, <given-names>V.</given-names></string-name>, &amp; <string-name><surname>Bannister</surname>, <given-names>K.</given-names></string-name> (<year>2005</year>). <article-title>Phonetic training with acoustic cue manipulations: A comparison of methods for teaching English /r/-/l/ to Japanese adults</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>118</volume>(<issue>5</issue>), <fpage>3267</fpage>&#8211;<lpage>3278</lpage>. <pub-id pub-id-type="doi">10.1121/1.2062307</pub-id></mixed-citation></ref>
<ref id="B38"><mixed-citation publication-type="journal"><string-name><surname>Kapnoula</surname>, <given-names>E. C.</given-names></string-name>, &amp; <string-name><surname>McMurray</surname>, <given-names>B.</given-names></string-name> (<year>2021</year>). <article-title>Idiosyncratic use of bottom-up and top-down information leads to differences in speech perception flexibility: Converging evidence from ERPs and eye-tracking</article-title>. <source>Brain and Language</source>, <volume>223</volume>, <elocation-id>105031</elocation-id>.</mixed-citation></ref>
<ref id="B39"><mixed-citation publication-type="journal"><string-name><surname>Kartushina</surname>, <given-names>N.</given-names></string-name>, <string-name><surname>Hervais-Adelman</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Frauenfelder</surname>, <given-names>U. H.</given-names></string-name>, &amp; <string-name><surname>Golestani</surname>, <given-names>N.</given-names></string-name> (<year>2016</year>). <article-title>Mutual influences between native and non-native vowels in production: Evidence from short-term visual articulatory feedback training</article-title>. <source>Journal of Phonetics</source>, <volume>57</volume>, <fpage>21</fpage>&#8211;<lpage>39</lpage>.</mixed-citation></ref>
<ref id="B40"><mixed-citation publication-type="webpage"><string-name><surname>Kimball</surname>, <given-names>A. E.</given-names></string-name> (<year>2014</year>). <article-title>The (statistical) power of Mechanical Turk. Purdue Linguistic Association Symposium</article-title>. <uri>https://docs.lib.purdue.edu/plas/2014/proceedings/1/</uri></mixed-citation></ref>
<ref id="B41"><mixed-citation publication-type="journal"><string-name><surname>Kimball</surname>, <given-names>A. E.</given-names></string-name>, <string-name><surname>Keupdjio</surname>, <given-names>H.</given-names></string-name>, <string-name><surname>Franich</surname>, <given-names>K.</given-names></string-name>, &amp; <string-name><surname>Kouankem</surname>, <given-names>C.</given-names></string-name> (<year>2019</year>). <article-title>Expanding field studies using online speech perception experiments</article-title>. In <source>Proceedings of the 19th International Congress of Phonetic Sciences</source> (pp. <fpage>315</fpage>&#8211;<lpage>319</lpage>).</mixed-citation></ref>
<ref id="B42"><mixed-citation publication-type="journal"><string-name><surname>Kirk</surname>, <given-names>N. W.</given-names></string-name> (<year>2023</year>). <article-title>MIND your language(s): Recognizing minority, indigenous, non-standard (ized), and dialect variety usage in &#8220;monolinguals.&#8221;</article-title> <source>Applied Psycholinguistics</source>, <volume>44</volume>(<issue>3</issue>), <fpage>358</fpage>&#8211;<lpage>364</lpage>.</mixed-citation></ref>
<ref id="B43"><mixed-citation publication-type="journal"><string-name><surname>Kostadinova</surname>, <given-names>V.</given-names></string-name>, &amp; <string-name><surname>Gardner</surname>, <given-names>M. H.</given-names></string-name> (<year>2024</year>, <month>January</month>). <article-title>Getting &#8220;good&#8221; data in a pandemic, part 1: Assessing the validity and quality of data collected remotely</article-title>. <source>Linguistics Vanguard</source>, <volume>9</volume>(<issue>s4</issue>), <fpage>329</fpage>&#8211;<lpage>334</lpage>. <pub-id pub-id-type="doi">10.1515/lingvan-2023-0170</pub-id></mixed-citation></ref>
<ref id="B44"><mixed-citation publication-type="webpage"><string-name><surname>Kunath</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Weinberger</surname>, <given-names>S. H.</given-names></string-name> (<year>2010</year>). <chapter-title>The wisdom of the crowd&#8217;s ear: Speech accent rating and annotation with Amazon Mechanical Turk</chapter-title>. In <source>Proceedings of the NAACL HLT 2010 workshop on creating speech and language data with Amazon&#8217;s Mechanical Turk</source> (pp. <fpage>168</fpage>&#8211;<lpage>171</lpage>). <uri>https://aclanthology.org/W10-0726.pdf</uri></mixed-citation></ref>
<ref id="B45"><mixed-citation publication-type="journal"><string-name><surname>Kutlu</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Hayes-Harb</surname>, <given-names>R.</given-names></string-name> (<year>2023</year>). <article-title>Towards a just and equitable applied psycholinguistics</article-title>. <source>Applied Psycholinguistics</source>, <volume>44</volume>(<issue>3</issue>), <fpage>293</fpage>&#8211;<lpage>300</lpage>.</mixed-citation></ref>
<ref id="B46"><mixed-citation publication-type="journal"><string-name><surname>Lively</surname>, <given-names>S. E.</given-names></string-name>, <string-name><surname>Logan</surname>, <given-names>J. S.</given-names></string-name>, &amp; <string-name><surname>Pisoni</surname>, <given-names>D. B.</given-names></string-name> (<year>1993</year>). <article-title>Training Japanese listeners to identify English /r/ and /l/. II: The role of phonetic environment and talker variability in learning new perceptual categories</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>94</volume>(<issue>3</issue>), <fpage>1242</fpage>&#8211;<lpage>1255</lpage>.</mixed-citation></ref>
<ref id="B47"><mixed-citation publication-type="journal"><string-name><surname>Logan</surname>, <given-names>J. S.</given-names></string-name>, <string-name><surname>Lively</surname>, <given-names>S. E.</given-names></string-name>, &amp; <string-name><surname>Pisoni</surname>, <given-names>D. B.</given-names></string-name> (<year>1991</year>). <article-title>Training Japanese listeners to identify English/r/and/l: A first report</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>89</volume>(<issue>2</issue>), <fpage>874</fpage>&#8211;<lpage>886</lpage>.</mixed-citation></ref>
<ref id="B48"><mixed-citation publication-type="journal"><string-name><surname>Mason</surname>, <given-names>W.</given-names></string-name>, &amp; <string-name><surname>Suri</surname>, <given-names>S.</given-names></string-name> (<year>2012</year>). <article-title>Conducting behavioral research on Amazon&#8217;s Mechanical Turk</article-title>. <source>Behavior Research Methods</source>, <volume>44</volume>(<issue>1</issue>), <fpage>1</fpage>&#8211;<lpage>23</lpage>. <pub-id pub-id-type="doi">10.3758/s13428-0110124-6</pub-id></mixed-citation></ref>
<ref id="B49"><mixed-citation publication-type="journal"><string-name><surname>Matthews</surname>, <given-names>G.</given-names></string-name>, &amp; <string-name><surname>Campbell</surname>, <given-names>S. E.</given-names></string-name> (<year>2009</year>). <article-title>Sustained performance under overload: Personality and individual differences in stress and coping</article-title>. <source>Theoretical Issues in Ergonomics Science</source>, <volume>10</volume>(<issue>5</issue>), <fpage>417</fpage>&#8211;<lpage>442</lpage>. <pub-id pub-id-type="doi">10.1080/14639220903106395</pub-id></mixed-citation></ref>
<ref id="B50"><mixed-citation publication-type="journal"><string-name><surname>McHaney</surname>, <given-names>J. R.</given-names></string-name>, <string-name><surname>Tessmer</surname>, <given-names>R.</given-names></string-name>, <string-name><surname>Roark</surname>, <given-names>C. L.</given-names></string-name>, &amp; <string-name><surname>Chandrasekaran</surname>, <given-names>B.</given-names></string-name> (<year>2021</year>). <article-title>Working memory relates to individual differences in speech category learning: Insights from computational modeling and pupillometry</article-title>. <source>Brain and Language</source>, <volume>222</volume>, <elocation-id>105010</elocation-id>.</mixed-citation></ref>
<ref id="B51"><mixed-citation publication-type="journal"><string-name><surname>McMurray</surname>, <given-names>B.</given-names></string-name> (<year>2022</year>). <article-title>The myth of categorical perception</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>152</volume>(<issue>6</issue>), <fpage>3819</fpage>&#8211;<lpage>3842</lpage>.</mixed-citation></ref>
<ref id="B52"><mixed-citation publication-type="journal"><string-name><surname>McMurray</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Baxelbaum</surname>, <given-names>K. S.</given-names></string-name>, <string-name><surname>Colby</surname>, <given-names>S.</given-names></string-name>, &amp; <string-name><surname>Tomblin</surname>, <given-names>J. B.</given-names></string-name> (<year>2023</year>). <article-title>Understanding language processing in variable populations on their own terms: Towards a functionalist psycholinguistics of individual differences, development, and disorders</article-title>. <source>Applied Psycholinguistics</source>, <volume>44</volume>(<issue>4</issue>), <fpage>565</fpage>&#8211;<lpage>592</lpage>.</mixed-citation></ref>
<ref id="B53"><mixed-citation publication-type="journal"><string-name><surname>Mora</surname>, <given-names>J. C.</given-names></string-name>, &amp; <string-name><surname>Darcy</surname>, <given-names>I.</given-names></string-name> (<year>2023</year>). <article-title>Individual differences in attention control and the processing of phonological contrasts in a second language</article-title>. <source>Phonetica</source>, <volume>80</volume>(<issue>3&#8211;4</issue>), <fpage>153</fpage>&#8211;<lpage>184</lpage>. <pub-id pub-id-type="doi">10.1515/phon-2022-0020</pub-id></mixed-citation></ref>
<ref id="B54"><mixed-citation publication-type="book"><string-name><surname>Mora</surname>, <given-names>J. C.</given-names></string-name>, &amp; <string-name><surname>Mora-Plaza</surname>, <given-names>I.</given-names></string-name> (<year>2019</year>). <chapter-title>Contributions of cognitive attention control to L2 speech learning</chapter-title>. In <string-name><given-names>A. M.</given-names> <surname>Nyvad</surname></string-name>, <string-name><given-names>M.</given-names> <surname>Hejn&#225;</surname></string-name>, <string-name><given-names>A.</given-names> <surname>H&#248;jen</surname></string-name>, <string-name><given-names>A. B.</given-names> <surname>Jespersen</surname></string-name>, &amp; <string-name><given-names>M. H.</given-names> <surname>S&#248;rensen</surname></string-name> (Eds.), <source>A sound approach to language matters&#8211;In honor of Ocke-Schwen Bohn</source>, <fpage>477</fpage>&#8211;<lpage>499</lpage>. <publisher-name>Dept. of English, School of Communication &amp; Culture, Aarhus University</publisher-name>, <publisher-loc>Denmark</publisher-loc>. <pub-id pub-id-type="doi">10.7146/aul.322.218</pub-id></mixed-citation></ref>
<ref id="B55"><mixed-citation publication-type="journal"><string-name><surname>Ortega</surname>, <given-names>L.</given-names></string-name> (<year>2005</year>). <article-title>For what and for whom is our research? The ethical as transformative lens in instructed SLA</article-title>. <source>Modern Language Journal</source>, <volume>89</volume>(<issue>3</issue>), <fpage>427</fpage>&#8211;<lpage>443</lpage>. <pub-id pub-id-type="doi">10.1111/j.1540-4781.2005.00315.x</pub-id></mixed-citation></ref>
<ref id="B56"><mixed-citation publication-type="journal"><string-name><surname>Pavlick</surname>, <given-names>E.</given-names></string-name>, <string-name><surname>Post</surname>, <given-names>M.</given-names></string-name>, <string-name><surname>Irvine</surname>, <given-names>A.</given-names></string-name>, <string-name><surname>Kachaev</surname>, <given-names>D.</given-names></string-name>, &amp; <string-name><surname>Callison-Burch</surname>, <given-names>C.</given-names></string-name> (<year>2014</year>). <article-title>The language demographics of Amazon Mechanical Turk</article-title>. <source>Transactions of the Association for Computational Linguistics</source>, <volume>2</volume>, <fpage>79</fpage>&#8211;<lpage>92</lpage>. <pub-id pub-id-type="doi">10.1162/tacl_a_00167</pub-id></mixed-citation></ref>
<ref id="B57"><mixed-citation publication-type="journal"><string-name><surname>Perrachione</surname>, <given-names>T. K.</given-names></string-name>, <string-name><surname>Lee</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Ha</surname>, <given-names>L. Y. Y.</given-names></string-name>, &amp; <string-name><surname>Wong</surname>, <given-names>P. C. M.</given-names></string-name> (<year>2011</year>). <article-title>Learning a novel phonological contrast depends on interactions between individual differences and training paradigm design</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>130</volume>(<issue>1</issue>), <elocation-id>461</elocation-id>. <pub-id pub-id-type="doi">10.1121/1.3593366</pub-id></mixed-citation></ref>
<ref id="B58"><mixed-citation publication-type="journal"><string-name><surname>Pisoni</surname>, <given-names>D. B.</given-names></string-name>, &amp; <string-name><surname>Tash</surname>, <given-names>J.</given-names></string-name> (<year>1974</year>). <article-title>Reaction times to comparisons within and across phonetic categories</article-title>. <source>Perception &amp; Psychophysics</source>, <volume>15</volume>(<issue>2</issue>), <fpage>285</fpage>&#8211;<lpage>290</lpage>.</mixed-citation></ref>
<ref id="B59"><mixed-citation publication-type="journal"><string-name><surname>Rad</surname>, <given-names>M. S.</given-names></string-name>, <string-name><surname>Martingano</surname>, <given-names>A. J.</given-names></string-name>, &amp; <string-name><surname>Ginges</surname>, <given-names>J.</given-names></string-name> (<year>2018</year>). <article-title>Toward a psychology of <italic>Homo sapiens</italic>: Making psychological science more representative of the human population</article-title>. <source>Proceedings of the National Academy of Sciences</source>, <volume>115</volume>(<issue>45</issue>), <fpage>11401</fpage>&#8211;<lpage>11405</lpage>. <pub-id pub-id-type="doi">10.1073/pnas.1721165115</pub-id></mixed-citation></ref>
<ref id="B60"><mixed-citation publication-type="journal"><string-name><surname>Roark</surname>, <given-names>C. L.</given-names></string-name>, <string-name><surname>Paulon</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>Rebaudo</surname>, <given-names>G.</given-names></string-name>, <string-name><surname>McHaney</surname>, <given-names>J. R.</given-names></string-name>, <string-name><surname>Sarkar</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Chandrasekaran</surname>, <given-names>B.</given-names></string-name> (<year>2022</year>). <article-title>Individual differences in working memory impact the trajectory of non-native speech category learning</article-title>. <source>PLOS One</source>, <volume>19</volume>(<issue>6</issue>), <elocation-id>e029717</elocation-id>. <pub-id pub-id-type="doi">10.1371/journal.pone.2097917</pub-id></mixed-citation></ref>
<ref id="B61"><mixed-citation publication-type="webpage"><string-name><surname>Sanker</surname>, <given-names>C.</given-names></string-name> (<year>2023</year>). <article-title>How do headphone checks impact perception data?</article-title> <source>Laboratory Phonology</source>, <volume>14</volume>(<issue>1</issue>). <pub-id pub-id-type="doi">10.16995/labphon.8778</pub-id></mixed-citation></ref>
<ref id="B62"><mixed-citation publication-type="journal"><string-name><surname>Schertz</surname>, <given-names>J.</given-names></string-name>, <string-name><surname>Cho</surname>, <given-names>T.</given-names></string-name>, <string-name><surname>Lotto</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Warner</surname>, <given-names>N.</given-names></string-name> (<year>2015</year>). <article-title>Individual differences in phonetic cue use in production and perception of a non-native sound contrast</article-title>. <source>Journal of Phonetics</source>, <volume>52</volume>, <fpage>183</fpage>&#8211;<lpage>204</lpage>. <pub-id pub-id-type="doi">10.1016/j.wocn.2015.07.003</pub-id></mixed-citation></ref>
<ref id="B63"><mixed-citation publication-type="journal"><string-name><surname>Schnoebelen</surname>, <given-names>T.</given-names></string-name>, &amp; <string-name><surname>Kuperman</surname>, <given-names>V.</given-names></string-name> (<year>2010</year>). <article-title>Using Amazon Mechanical Turk for linguistic research</article-title>. <source>Psihologija</source>, <volume>43</volume>(<issue>4</issue>), <fpage>441</fpage>&#8211;<lpage>464</lpage>. <pub-id pub-id-type="doi">10.2298/PSI1004441S</pub-id></mixed-citation></ref>
<ref id="B64"><mixed-citation publication-type="journal"><string-name><surname>Schouten</surname>, <given-names>B.</given-names></string-name>, <string-name><surname>Gerrits</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Van Hessen</surname>, <given-names>A.</given-names></string-name> (<year>2003</year>). <article-title>The end of categorical perception as we know it</article-title>. <source>Speech Communication</source>, <volume>41</volume>(<issue>1</issue>), <fpage>71</fpage>&#8211;<lpage>80</lpage>.</mixed-citation></ref>
<ref id="B65"><mixed-citation publication-type="journal"><string-name><surname>Staggs</surname>, <given-names>C.</given-names></string-name>, <string-name><surname>Baese-Berk</surname>, <given-names>M. M.</given-names></string-name>, &amp; <string-name><surname>Nagle</surname>, <given-names>C.</given-names></string-name> (<year>2022</year>). <article-title>The influence of social information on speech intelligibility within the Spanish Heritage community</article-title>. <source>Languages</source>, <volume>7</volume>(<issue>3</issue>), <elocation-id>231</elocation-id>. <pub-id pub-id-type="doi">10.3390/languages7030231</pub-id></mixed-citation></ref>
<ref id="B66"><mixed-citation publication-type="journal"><string-name><surname>Stevens</surname>, <given-names>K. N.</given-names></string-name>, &amp; <string-name><surname>Blumstein</surname>, <given-names>S. E.</given-names></string-name> (<year>1975</year>). <article-title>Quantal aspects of consonant production and perception: A study of retroflex stop consonants</article-title>. <source>Journal of Phonetics</source>, <volume>3</volume>(<issue>4</issue>), <fpage>215</fpage>&#8211;<lpage>233</lpage>. <pub-id pub-id-type="doi">10.1016/S0095-4470(19)31431-7</pub-id></mixed-citation></ref>
<ref id="B67"><mixed-citation publication-type="journal"><string-name><surname>Strand</surname>, <given-names>J. F.</given-names></string-name>, <string-name><surname>Brown</surname>, <given-names>V. A.</given-names></string-name>, <string-name><surname>Sewell</surname>, <given-names>K.</given-names></string-name>, <string-name><surname>Lin</surname>, <given-names>Y.</given-names></string-name>, <string-name><surname>Lefkowitz</surname>, <given-names>E.</given-names></string-name>, &amp; <string-name><surname>Saksena</surname>, <given-names>C. G.</given-names></string-name> (<year>2024</year>). <article-title>Assessing the effects of &#8220;native speaker&#8221; status on classic findings in speech research</article-title>. <source>Journal of Experimental Psychology: General</source>, <volume>153</volume>(<issue>12</issue>), <fpage>3027</fpage>&#8211;<lpage>3041</lpage>. <pub-id pub-id-type="doi">10.1037/xge0001640</pub-id></mixed-citation></ref>
<ref id="B68"><mixed-citation publication-type="journal"><string-name><surname>Strange</surname>, <given-names>W.</given-names></string-name>, &amp; <string-name><surname>Dittman</surname>, <given-names>S.</given-names></string-name> (<year>1984</year>). <article-title>Effects of discrimination training on the perception of /r-l/ by Japanese adults learning English</article-title>. <source>Perception and Psychophysics</source>, <volume>36</volume>, <fpage>131</fpage>&#8211;<lpage>145</lpage>.</mixed-citation></ref>
<ref id="B69"><mixed-citation publication-type="journal"><string-name><surname>Tripp</surname>, <given-names>A.</given-names></string-name>, &amp; <string-name><surname>Munson</surname>, <given-names>B.</given-names></string-name> (<year>2023</year>). <article-title>Acknowledging language variation and its power: Keys to justice and equity in applied psycholinguistics</article-title>. <source>Applied Psycholinguistics</source>, <volume>44</volume>(<issue>4</issue>), <fpage>495</fpage>&#8211;<lpage>513</lpage>. <pub-id pub-id-type="doi">10.1017/S0142716423000206</pub-id></mixed-citation></ref>
<ref id="B70"><mixed-citation publication-type="journal"><string-name><surname>West</surname>, <given-names>R.</given-names></string-name> (<year>1999</year>). <article-title>Visual distraction, working memory, and aging</article-title>. <source>Memory &amp; Cognition</source>, <volume>27</volume>(<issue>6</issue>), <fpage>1064</fpage>&#8211;<lpage>1072</lpage>. <pub-id pub-id-type="doi">10.3758BF03201235</pub-id></mixed-citation></ref>
<ref id="B71"><mixed-citation publication-type="journal"><string-name><surname>Woods</surname>, <given-names>K. J. P.</given-names></string-name>, <string-name><surname>Siegel</surname>, <given-names>M. H.</given-names></string-name>, <string-name><surname>Traer</surname>, <given-names>J.</given-names></string-name>, &amp; <string-name><surname>McDermott</surname>, <given-names>J. H.</given-names></string-name> (<year>2017</year>). <article-title>Headphone screening to facilitate web-based auditory experiments</article-title>. <source>Attention, Perception, and Psychophysics</source>, <volume>79</volume>(<issue>7</issue>), <fpage>2064</fpage>&#8211;<lpage>2072</lpage>.</mixed-citation></ref>
<ref id="B72"><mixed-citation publication-type="journal"><string-name><surname>Yu</surname>, <given-names>A. C.</given-names></string-name>, &amp; <string-name><surname>Lee</surname>, <given-names>H.</given-names></string-name> (<year>2014</year>). <article-title>The stability of perceptual compensation for coarticulation within and across individuals: A cross-validation study</article-title>. <source>Journal of the Acoustical Society of America</source>, <volume>136</volume>(<issue>1</issue>), <fpage>382</fpage>&#8211;<lpage>388</lpage>. <pub-id pub-id-type="doi">10.1121/1.4883380</pub-id></mixed-citation></ref>
</ref-list>
</back>
</article>