<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "journalpublishing3.dtd">
<article article-type="research-article" dtd-version="3.0" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">LOQ</journal-id>
<journal-title-group>
<journal-title>Loquens</journal-title>
</journal-title-group>
<issn pub-type="epub">2386-2637</issn>
<publisher>
<publisher-name>Consejo Superior de Investigaciones Cientificas</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">loquens_060</article-id>
<article-id pub-id-type="doi">10.3989/loquens.2019.061</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Articles</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Is there an interlanguage intelligibility benefit in perception of English word stress?</article-title>
<trans-title-group xml:lang="es">
<trans-title>&#x00BF;Existe un beneficio de inteligibilidad por interlengua en la percepci&#x00F3;n del acento t&#x00F3;nico?&#x2013;</trans-title>
</trans-title-group>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Almbark</surname>
<given-names>Rana</given-names>
</name>
<xref ref-type="aff" rid="aff0001">1</xref>
<xref ref-type="aff" rid="aff0002">2</xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Bouchhioua</surname>
<given-names>Nadia</given-names>
</name>
<xref ref-type="aff" rid="aff0003">3</xref>
</contrib>
<contrib contrib-type="author" corresp="yes">
<name>
<surname>Hellmuth</surname>
<given-names>Sam</given-names>
</name>
<xref ref-type="aff" rid="aff0001">1</xref>
</contrib>
</contrib-group>
<aff id="aff0001"><label>1</label>University of York, United Kingdom</aff>
<aff id="aff0002"><label>2</label>University of Chester, United Kingdom</aff>
<aff id="aff0003"><label>3</label>Universit&#x00E9; de la Manouba, Tunisia</aff>
<author-notes>
<corresp id="cor1"><email xlink:href="rana.alhusseinalmbark@york.ac.uk">rana.alhusseinalmbark@york.ac.uk</email> ORCID: <ext-link ext-link-type="uri" xlink:href="https://orcid.org/0000-0002-4784-2497">https://orcid.org/0000-0002-4784-2497</ext-link>, <email xlink:href="Nadia.Bouchhioua@flm.rnu.tn">Nadia.Bouchhioua@flm.rnu.tn</email> ORCID: <ext-link ext-link-type="uri" xlink:href="https://orcid.org/0000-0001-7602-797X">https://orcid.org/0000-0001-7602-797X</ext-link>, <email xlink:href="sam.hellmuth@york.ac.uk">sam.hellmuth@york.ac.uk</email> ORCID: <ext-link ext-link-type="uri" xlink:href="https://orcid.org/0000-0002-0062-904X">https://orcid.org/0000-0002-0062-904X</ext-link></corresp>
</author-notes>
<pub-date pub-type="epub">
<day>31</day>
<month>01</month>
<year>2019</year>
</pub-date>
<pub-date pub-type="collection">
<year>2019</year>
</pub-date>
<volume>06</volume>
<issue>1</issue>
<elocation-id content-type="doi">10.3989/loquens.2019.061</elocation-id>
<history>
<date date-type="received">
<day>05</day>
<month>03</month>
<year>2019</year>
</date>
<date date-type="accepted">
<day>14</day>
<month>03</month>
<year>2019</year>
</date>
<date date-type="published online">
<day>22</day>
<month>07</month>
<year>2019</year>
</date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2019 CSIC</copyright-statement>
<copyright-year>2019</copyright-year>
<license license-type="open-access" xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution 4.0 International (CC BY 4.0) License.</license-p>
</license>
</permissions>
<abstract>
<title>ABSTRACT</title>
<p>This paper asks whether there is an &#x2018;interlanguage intelligibility benefit&#x2019; in perception of word-stress, as has been reported for global sentence recognition. L1 English listeners, and L2 English listeners who are L1 speakers of Arabic dialects from Jordan and Egypt, performed a binary forced-choice identification task on English near-minimal pairs (such as[&#x02C8;&#x0252;bd&#x0292;&#x025B;kt] ~ [&#x0259;b&#x02C8;d&#x0292;&#x025B;kt]) produced by an L1 English speaker, and two L2 English speakers from Jordan and Egypt respectively. The results show an overall advantage for L1 English listeners, which replicates the findings of an earlier study for general sentence recognition, and which is also consistent with earlier findings that L1 listeners rely more on structural knowledge than on acoustic cues in stress perception. Non-target-like L2 productions of words with final stress (which are primarily cued in L1 production by vowel reduction in the initial unstressed syllable) were less accurately recognized by L1 English listeners than by L2 listeners, but there was no evidence of a generalized advantage for L2 listeners in response to other L2 stimuli.</p>
</abstract>
<trans-abstract xml:lang="es">
<title>RESUMEN</title>
<p><italic>&#x00BF;Existe un beneficio de inteligibilidad por interlengua en la percepci&#x00F3;n del acento t&#x00F3;nico?&#x2013;.</italic> Este trabajo pregunta si existe un &#x201C;beneficio de inteligibilidad por interlengua&#x201D; (&#x2018;interlanguage intelligibility benefit&#x2019;) en la percepci&#x00F3;n del acento t&#x00F3;nico, como se ha reportado para el reconocimiento global de oraciones. El estudio involucr&#x00F3; a un grupo de oyentes de ingl&#x00E9;s como primera lengua (L1), y a oyentes nativos de dialectos &#x00E1;rabes de Jordania y Egipto que utilizan ingl&#x00E9;s como segunda lengua (L2), quienes participaron como jueces perceptivos en una tarea de identificaci&#x00F3;n de respuesta binaria forzada. El est&#x00ED;mulo estuvo conformado por pares casi m&#x00ED;nimos del ingl&#x00E9;s (por ejemplo, [&#x02C8;&#x0252;bd&#x0292;&#x025B;kt] ~ [&#x0259;b&#x02C8;d&#x0292;&#x025B;kt]) producidos por un hablante nativo de ingl&#x00E9;s, y por dos hablantes de ingl&#x00E9;s como L2 de dialectos de Jordania y Egipto, respectivamente. Los resultados revelan una ventaja para los oyentes nativos de ingl&#x00E9;s, lo que replica los resultados de un estudio previo sobre reconocimiento de oraciones, y tambi&#x00E9;n es consistente con descubrimientos anteriores que especificaron que los oyentes nativos utilizan mayormente conocimientos estructurales en la percepci&#x00F3;n del acento tonal en lugar de los marcadores de la se&#x00F1;al ac&#x00FA;stica. Realizaciones no nativas de las palabras acentuadas en la &#x00FA;ltima s&#x00ED;laba (que est&#x00E1;n marcadas por una reducci&#x00F3;n voc&#x00E1;lica en la primera s&#x00ED;laba en ingl&#x00E9;s nativo) fueron reconocidas menos exitosamente por los oyentes nativos en ingl&#x00E9;s, pero no hubo evidencia de una ventaja generalizada para los oyentes en un segundo idioma cuando escuchan a otros hablantes del mismo primer idioma.</p></trans-abstract>
<kwd-group xml:lang="en">
<title>Keywords</title>
<kwd>interlanguage intelligibility benefit</kwd>
<kwd>word-stress</kwd>
<kwd>perception</kwd>
<kwd>L2 English</kwd>
<kwd>L1 Arabic</kwd>
</kwd-group>
<kwd-group xml:lang="es">
<title>Palabras clave</title>
<kwd>beneficio de inteligibilidad por interlengua</kwd>
<kwd>acento t&#x00F3;nico</kwd>
<kwd>percepci&#x00F3;n</kwd>
<kwd>ingl&#x00E9;s como L2</kwd>
<kwd>&#x00E1;rabe como L1</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="sec1" sec-type="intro">
<title>1. INTRODUCTION</title>
<p>An &#x2018;interlanguage intelligibility benefit&#x2019; has been reported for global sentence perception (Bent &#x0026; Bradlow, <xref ref-type="bibr" rid="cit0004">2003</xref>), whereby L2 English listeners outperform L1 English listeners in a sentence recognition task on the productions of other L2 speakers. In the present paper we explore whether a similar effect holds in the narrow domain of L2 listeners&#x2019; perception of English word-stress. Specifically, we explore whether non-target-like phonetic realization of stress in L2 speakers&#x2019; productions results in intelligibility issues for L1 and/or L2 listeners in a word recognition task on English stress near-minimal pairs. We use speech stimuli extracted from larger utterances elicited using a carefully controlled paradigm so that the cues to stress in the stimuli are those to word-level stress only, without any enhancement due to phrase- or sentence-level prominence. The present study thus offers a first exploration of an eventual interlanguage intelligibility benefit due to transfer of L1 patterns in the acoustic realization of stress into L2 productions. We also explore the general issue of whether non-target-like acoustic realization of word stress leads to reduced intelligibility of L2 speech, by L1 and/or L2 listeners.</p>
<p>We use the term &#x2018;stress&#x2019; to denote word-level stress or lexical prominence, and the term &#x2018;accent&#x2019; to denote phrase-level stress or post-lexical prominence. The focus of our study is word-level stress as produced and perceived by speakers of English as first (L1) and as second or additional (L2) language. We note that&#x2014;to investigate stress in languages such as English and Arabic in which both stress and accent are marked (Jun, <xref ref-type="bibr" rid="cit0020">2014</xref>)&#x2014;it is necessary to control for the presence or absence of accent (Beckman &#x0026; Edwards, <xref ref-type="bibr" rid="cit0003">1994</xref>; Roettger &#x0026; Gordon, <xref ref-type="bibr" rid="cit0026">2017</xref>).</p>
</sec>
<sec id="sec2">
<title>2. BACKGROUND TO THE STUDY</title>
<sec id="sec2.1">
<title>2.1. The correlates of stress in production</title>
<p>The acoustic correlates of stress have been shown to include duration, F0, overall intensity, frequency-sensitive intensity (spectral balance) and formant frequencies (F1/F2). Gordon and Roettger (<xref ref-type="bibr" rid="cit0018">2017</xref>) surveyed 110 studies on 75 languages and found that although duration was the most frequently observed cue to stress, all of these cues played a role of some kind in most of the languages surveyed. The relative strength of different cues appears to vary across languages, however.</p>
<p>It is widely assumed that F0 is the most prominent and consistent cue to stress in English, based on the influential early study by Fry (<xref ref-type="bibr" rid="cit0017">1955</xref>), which did not, however, examine the correlates of stress in the absence of accent. Studies which avoid the stress versus accent confound instead report duration, spectral balance and formant frequencies as the most consistent cues in English (Bouchhioua, <xref ref-type="bibr" rid="cit0006">2016</xref>; van Heuven &#x0026; Sluijter, <xref ref-type="bibr" rid="cit0030">1996</xref>).</p>
<p>There has been less prior investigation of the acoustic correlates of stress in production of Arabic stress. Cross-dialectal variation in the acoustic cues to stress is likely, since cross-dialectal variation in phonological stress assignment is well established (Watson, <xref ref-type="bibr" rid="cit0032">2011</xref>). In addition, some dialects such as Egyptian Arabic (EA) display consistent co-occurrence of stress and accent: the stressed syllable of almost all content words also carries sentence-level accent (Chahal &#x0026; Hellmuth, <xref ref-type="bibr" rid="cit0007">2015</xref>).</p>
<p>One of the first studies of the correlates of stress in Arabic was on Jordanian Arabic (JA), and indicated that the cues to stress in JA are duration and F1 (de Jong &#x0026; Zawaydeh, <xref ref-type="bibr" rid="cit0013">1999</xref>). In contrast, the correlates of stress reported for Tunisian Arabic are spectral balance and F1, but not duration (Bouchhioua, <xref ref-type="bibr" rid="cit0006">2016</xref>).</p>
<p>In a previous study we compared the correlates of stress in JA and EA&#x2014;the two dialects investigated in the present study&#x2014;and found that both dialects made use of duration, intensity and F0, but not formant frequencies or spectral balance (Almbark, Bouchhioua, &#x0026; Hellmuth, <xref ref-type="bibr" rid="cit0001">2014</xref>). The only differences between JA and EA were in the degree to which cues were used: there was greater differentiation of stressed and unstressed syllables by means of duration in EA than in JA, and by means of F0 in JA than in EA. This finding for JA contrasts with that of the earlier study of JA by de Jong and Zawaydeh (<xref ref-type="bibr" rid="cit0013">1999</xref>), which did not fully control for the confound of stress and accent.</p>
</sec>
<sec id="sec2.2">
<title>2.2. Perception of the correlates of stress</title>
<p>There is also cross-linguistic variation in the relative weighting of acoustic cues to stress in perception, and in the extent to which acoustic cues are relied upon compared to other factors.</p>
<p>Several studies have shown that listeners may rely on only a subset of the available acoustic cues in the signal. A recent study explored the perceptual behavior of English, Russian and Mandarin listeners in a forced choice identification task, in response to disyllabic pseudo-word stimuli in which F0, duration, intensity and F1/F2 of target vowels was systematically varied; vowel quality (F1/F2) had the greatest influence on the choices of listeners from all three language backgrounds, but there was variation in the relative weighting of suprasegmental cues (Chrabaszcz, Winn, Lin, &#x0026; Idsardi, <xref ref-type="bibr" rid="cit0008">2014</xref>). F0 was the next strongest cue after F1/F2 for English and Mandarin listeners, but duration and intensity were more important for Russian listeners. Similarly, Standard Mandarin listeners are influenced in their perception of stress minimal pairs, in a sequence recall task, by both duration and F0 cues; this contrasts with Taiwanese Mandarin listeners who attend primarily to F0, reflecting the lack of use of durational cues to word-level prominence asymmetries in Taiwanese Mandarin (Qin, Chien, &#x0026; Tremblay, <xref ref-type="bibr" rid="cit0025">2017</xref>). In lexical retrieval tasks, English listeners in fact rely primarily on segmental cues provided by unstressed vowel reduction: the true minimal pair &#x2018;forebear&#x2019; (n.) [&#x02C8;f&#x0254;&#x02D0;b&#x025B;&#x0259;] ~ &#x2018;forebear&#x2019; (v.) [f&#x0254;&#x02D0;&#x02C8;b&#x025B;&#x0259;]&#x2014;in which there are no segmental cues to stress in the form of vowel reduction in the unstressed syllable&#x2014;is homophonous in perception for English listeners (Cutler, <xref ref-type="bibr" rid="cit0012">1986</xref>).</p>
<p>Stress perception is also influenced by the phonological status of stress in the listener&#x2019;s first language (L1). French is a language which does not display word-level stress, and a sequence of studies has shown that although French listeners are able to perceive the acoustic cues to stress in an AX discrimination task, they are unable to discriminate stress minimal pairs in a sequence recall task which requires phonological encoding of those acoustic cues in lexical representations (Dupoux, Pallier, Sebasti&#x00E1;n-Gall&#x00E9;s, &#x0026; Mehler, <xref ref-type="bibr" rid="cit0014">1997</xref>); this holds even after long-term exposure to (and advanced proficiency in) Spanish, which is a language with contrastive stress (Dupoux, Sebasti&#x00E1;n-Gall&#x00E9;s, Navarrete, &#x0026; Peperkamp, <xref ref-type="bibr" rid="cit0015">2008</xref>).</p>
<p>Finally, perception of stress is not influenced solely by acoustic correlates to stress and their relative weighting or phonological status. Several studies have shown that &#x2018;bottom-up&#x2019; phonetic cues are used alongside &#x2018;top-down&#x2019; cues such as lexico-semantic information in perception and processing of stress (Cole, Mo, &#x0026; Hasegawa-Johnson, <xref ref-type="bibr" rid="cit0009">2010</xref>; Eriksson, Thunberg, &#x0026; Traunm&#x00FC;ller, <xref ref-type="bibr" rid="cit0016">2001</xref>). Mattys, White, and Melhorn (<xref ref-type="bibr" rid="cit0023">2005</xref>) argue that English listeners rely on different types of cues in a word segmentation task, with cues forming a hierarchy: fine-grained phonetic cues to stress are argued to be lower in the hierarchy than lexical and semantic cues, because phonetic cues are only relied on when performing the task in adverse listening conditions. This may be one strategy which allows listeners to use &#x2018;perceptual normalization&#x2019; to recover the hypothesized intended form from non-target-like realizations (Ohala, <xref ref-type="bibr" rid="cit0024">1993</xref>). In contrast, L2 listeners show less reliance than L1 listeners on &#x2018;top-down&#x2019; structural or lexical information in a word-by-word prominence rating task; instead, L2 listeners&#x2019; ratings more closely reflected differences in the relative strength of acoustic phonetic cues (Wagner, <xref ref-type="bibr" rid="cit0031">2005</xref>).</p>
</sec>
<sec id="sec2.3">
<title>2.3. The interlanguage intelligibility benefit</title>
<p>The term &#x2018;interlanguage&#x2019; describes patterns of language use, displayed by second language learners, which fall somewhere between the grammar of the native language and the target language being acquired (Selinker, <xref ref-type="bibr" rid="cit0027">1972</xref>).</p>
<p>The concept of an interlanguage speech intelligibility benefit was proposed by Bent and Bradlow (<xref ref-type="bibr" rid="cit0004">2003</xref>) to explain their findings in a sentence recognition task performed on L1 and L2 English speech samples, by L1 English listeners in comparison to L2 English listeners whose L1 varied. For native English listeners, the native English speech was more intelligible (more keywords accurately recognized) than the L2 English speech; however, for the L2 English listeners, the L1 English and L2 English speech were equally intelligible, regardless of whether the L2 English listener&#x2019;s L1 background matched that of the L2 English speaker they were listening to.</p>
<p>The two main groups of L2 English listeners in the Bent and Bradlow (<xref ref-type="bibr" rid="cit0004">2003</xref>) study were L1 speakers of Chinese and Korean. Stibbard and Lee (<xref ref-type="bibr" rid="cit0028">2006</xref>) replicated the same study design with L2 English speakers/listeners from more typologically diverse L1 backgrounds, however, and obtained a more nuanced result. They explored the perceptual behavior of L2 English speakers from Saudi Arabia or Korea, at two proficiency levels in English (low and high). In their study, the L1 English listener group showed higher recognition rates than any of the L2 listener groups, but high proficiency L2 English samples were equally well recognized as L1 English samples by both L1 and L2 listeners. The main finding of the replication study was that low proficiency was highly correlated with low intelligibility, as might be expected, but also that there was a matched interlanguage speech intelligibility benefit: low proficiency L2 English speech was better recognized by L2 listeners from the same L1 background as the speaker in the L2 English sample.</p>
<p>In this study we explore whether there is an interlanguage speech intelligibility benefit in respect of L1 versus L2 realization of the phonetic cues to word-level stress.</p>
</sec>
<sec id="sec2.4">
<title>2.4. The present study</title>
<p>The main research question of the paper is to determine whether there is an interlanguage intelligibility benefit in perception of English word stress. We use stimuli that were elicited using a paradigm designed to elicit English stress near-minimal pairs in a context in which the target word is realized without a phrase-level accent, thus focusing on listeners&#x2019; ability to make use of the phonetic cues to stress in the absence of cues to accent. Since vowel reduction is the primary cue to word stress for native English listeners (as noted in 2.2 above), it was important to use stimuli in which vowel reduction could appear, to determine whether failure to produce target words with appropriate vowel reduction reduces intelligibility, and perhaps differentially so for native versus non-native listeners. We therefore used near-minimal pairs in which vowel reduction in the unstressed syllable provides a segmental cue to stress alongside suprasegmental cues such as duration and intensity. The stimuli were produced by an L1 English native speaker (NE) and two L2 English non-native speakers (L2) from Jordan and Egypt, respectively. The listeners in a forced-choice identification task are L1 English listeners (NE) and L2 English listeners from Jordan and Egypt (L2). The over-arching research question stated in the title of this paper thus breaks down into three sub-questions, which we address in the present study by exploring the interaction of listener language and stimulus language in a single study with a crossed factor design:</p>
<list list-type="order">
<list-item>
<p>Do NE listeners identify the position of stress in the productions of a NE speaker more accurately than in those of L2 speakers?</p>
</list-item>
<list-item>
<p>Do L2 listeners identify the position of stress in the productions of L2 speakers more accurately than in those of a NE speaker?</p>
</list-item>
<list-item>
<p>Do L2 listeners identify the position of stress in the productions of an L2 speaker from their own L1 dialect background more accurately than in those of an L2 speaker from a different L1 dialect background?</p>
</list-item>
</list>
<p>Based on Bent and Bradlow&#x2019;s (<xref ref-type="bibr" rid="cit0004">2003</xref>) findings, we would predict an advantage for NE listeners when listening to NE productions, but no advantage for L2 listeners when listening to other L2 listeners (from any background). Based on Stibbard and Lee&#x2019;s (<xref ref-type="bibr" rid="cit0028">2006</xref>) findings, however, we predict an overall advantage for NE listeners, but a possible advantage for L2 listeners when listening to L2 listeners. Our interpretation of the results will also consider whether there are differences between NE and L2 listeners in reliance on &#x2018;bottom-up&#x2019; phonetic cues versus &#x2018;top-down&#x2019; structure-based expectations, by examining possible transfer effects which reflect the different structural properties of stress assignment in listeners&#x2019; L1.</p>
</sec>
</sec>
<sec id="sec3" sec-type="methods">
<title>3. METHODS</title>
<sec id="sec3.1">
<title>3.1. Materials</title>
<p>Stimuli which contrast in the position of stress were elicited using the nine English disyllabic near-minimal pairs, listed in <xref ref-type="table" rid="t0001">Table 1</xref>, following Bouchhioua (<xref ref-type="bibr" rid="cit0005">2008</xref>, <xref ref-type="bibr" rid="cit0006">2016</xref>).</p>
<table-wrap id="t0001">
<label>Table 1</label>
<caption>
<p>English near-minimal pairs, with stress on the first or second syllable.</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th colspan="2" align="center">stress on first syllable</th>
<th colspan="2" align="center">stress on second syllable</th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">&#x02C8;s&#x028C;bd&#x0292;&#x025B;kt</td>
<td align="left">subject (n.)</td>
<td align="left">s&#x0259;b&#x02C8;d&#x0292;&#x025B;kt</td>
<td align="left">subject (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;&#x0279;&#x025B;k&#x0254;&#x02D0;d</td>
<td align="left">record (n.)</td>
<td align="left">&#x0279;&#x026A;&#x02C8;k&#x0254;&#x02D0;d</td>
<td align="left">record (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;k&#x0252;nt&#x0279;&#x00E6;st</td>
<td align="left">contrast (n.)</td>
<td align="left">k&#x0259;n&#x02C8;t&#x0279;&#x00E6;st</td>
<td align="left">contrast (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;da&#x026A;d&#x0292;&#x025B;st</td>
<td align="left">digest (n.)</td>
<td align="left">d&#x026A;&#x02C8;d&#x0292;&#x025B;st</td>
<td align="left">digest (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;k&#x0252;nt&#x0279;&#x00E6;kt</td>
<td align="left">contract (n.)</td>
<td align="left">k&#x0259;n&#x02C8;t&#x0279;&#x00E6;kt</td>
<td align="left">contract (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;p&#x025C;&#x02D0;m&#x026A;t</td>
<td align="left">permit (n.)</td>
<td align="left">p&#x0259;&#x02C8;m&#x026A;t</td>
<td align="left">permit (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;&#x0252;bd&#x0292;&#x025B;kt</td>
<td align="left">object (n.)</td>
<td align="left">&#x0259;b&#x02C8;d&#x0292;&#x025B;kt</td>
<td align="left">object (v.)</td>
</tr>
<tr>
<td align="left">&#x02C8;k&#x0252;nt&#x025B;nt</td>
<td align="left">content (n.)</td>
<td align="left">k&#x0259;n&#x02C8;t&#x025B;nt</td>
<td align="left">content (adj.)</td>
</tr>
<tr>
<td align="left">&#x02C8;k&#x0252;nd&#x028C;kt</td>
<td align="left">conduct (n.)</td>
<td align="left">k&#x0259;n&#x02C8;d&#x028C;kt</td>
<td align="left">conduct (v.)</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>Three further pairs (<italic>combine</italic>, <italic>pervert</italic>, and <italic>project</italic>) were recorded but later excluded from the study, as stress was frequently misplaced due to unfamiliarity with the word in one or both stress positions. Six target-like tokens of the word <italic>project</italic> (two from each speaker) were used for the training phase of the experiment as outlined further below.</p>
<p>The intended accent status of the target word was varied by using a carrier phrase that either attracts focus to the target word [+accent] or diverts focus away from it [&#x2212;accent], again following Bouchhioua (<xref ref-type="bibr" rid="cit0005">2008</xref>, <xref ref-type="bibr" rid="cit0006">2016</xref>), as shown in <xref ref-type="table" rid="t0002">Table 2</xref>. The target word was always elicited in a carrier phrase: &#x2018;say ___ again&#x2019;. To attract accent onto the target word, a semantically related word preceded the target word in the same carrier phrase. To divert focus away from the target word, two preceding sentences are used to ensure that the target word appears in post-focal position (after the contrastively focused verb &#x2018;SAY&#x2019;) and is interpreted as old information due to being repeated from the immediately preceding discourse (Cruttenden, <xref ref-type="bibr" rid="cit0011">2006</xref>; Ladd, <xref ref-type="bibr" rid="cit0021">2008</xref>). Each sentence ~ context combination was read aloud once; sentences were presented to participants in pseudo-random order on a printed sheet.</p>
<table-wrap id="t0002">
<label>Table 2</label>
<caption>
<p>Target word (in bold) placed in carrier phrases to vary &#x00B1;accent status. CAPITALS denote expected position of sentence accents under focus.</p>
</caption>
<table frame="hsides" rules="groups">
<tbody>
<tr>
<td align="left">+accent</td>
<td align="left">Say topic again.<break/>Say SUBJECT again.</td>
</tr>
<tr>
<td align="left">&#x2212;accent</td>
<td align="left">The subject is a grammatical category.<break/>WRITE subject again.<break/>SSAY <bold>subject</bold> again. &#x2190;</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The experimental stimuli for the present study were extracted from target-like tokens (as judged to consensus by the first and third authors) produced in &#x2212;<italic>accent</italic> condition, as in (1), to investigate the extent to which listeners were able to detect phonetic cues to stress produced by the speakers, in the absence of any additional cues to accent.</p>
<table-wrap id="ut0001">
<table frame="void" rules="none">
<tbody>
<tr>
<td align="left">(1)</td>
<td align="left">stress on<break/>first syllable:</td>
<td align="left">SAY &#x02C8;s&#x028C;bd&#x0292;&#x025B;kt again.</td>
</tr>
<tr>
<td align="left"/>
<td align="left">stress on <break/>second syllable:</td>
<td align="left">SAY s&#x0259;b&#x02C8;d&#x0292;&#x025B;kt again</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>The stimuli for the perception experiment were produced by three male speakers, from: Cairo, Egypt (EA); Amman, Jordan (JA); UK (native speaker of British English, NE). The speakers were aged 26, 20, and 39 years, respectively. The Arabic speakers had learned English at school for 12 years but had never resided in an English-speaking country; they were selected from participants in an earlier production study (Almbark <italic>et al.</italic>, <xref ref-type="bibr" rid="cit0001">2014</xref>). Recordings were made in Cairo, Amman and York, respectively. Recordings were made in .wav format at 44.1 KHz 16 bit, on a Marantz PMD660 with external Shure SM10 headset microphone.</p>
<p>The results of acoustic analysis of the selected stimuli for duration, F0, intensity, F1/F2 and two measures of spectral tilt (H1.H2 or H1.A3), comparing properties of the vowel in the initial syllable (only), in stressed and unstressed condition, are illustrated in <xref ref-type="fig" rid="f0001">Figures 1</xref>&#x2013;<xref ref-type="fig" rid="f0002">2</xref>. We used a normalized vowel duration measure to control for inter-speaker variation in speech rate, by calculating vowel duration as a proportion of the whole word. The acoustic properties of the stimuli were explored in a series of linear mixed models (LMM) using <italic>lme4</italic> (Bates, Maechler, Bolker, &#x0026; Walker, <xref ref-type="bibr" rid="cit0002">2015</xref>) in <italic>R</italic> (Core Team, <xref ref-type="bibr" rid="cit0010">2014</xref>), with each acoustic measure in turn as dependent variable, <italic>speaker</italic> (EA ~ JA ~ NE) and <italic>stress</italic> (stressed ~ unstressed) and their interaction as fixed factors, and a random intercept for <italic>item</italic>.</p>
<fig id="f0001">
<label>Figure 1</label>
<caption>
<p>Median and interquartile range for values of (from left to right) maximum F0, peak intensity, normalized vowel duration and two measures of spectral emphasis H1&#x2013;H2 and H1A3, in the first vowel of experimental stimuli produced by the native speaker of Egyptian Arabic (EA; top row), Jordanian Arabic (JA; middle row), and English native speaker (NE; bottom row), grouped by stress condition (whether the vowel in which measurements was taken was stressed or unstressed).</p>
</caption>
<graphic xlink:href="loquens_061-g001.tif" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</fig>
<fig id="f0002">
<label>Figure 2</label>
<caption>
<p>F1/F2 plot of the first vowel in experimental stimuli produced by the native speaker of Egyptian Arabic (EA), Jordanian Arabic (JA) and English (NE), where the vowel is stressed (black dots) or unstressed (white dots).</p>
</caption>
<graphic xlink:href="loquens_061-g002.tif" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</fig>
<p>The acoustic analysis shows that F0 differentiates stressed and unstressed syllables in the L2 English productions of the JA speaker, but not in those of the EA or NE speakers. Similarly, although intensity is somewhat higher in stressed syllables than unstressed syllables for all three speakers, including the EA speaker (for whom intensity is the strongest cue to stress on average), nevertheless it is only in the JA speaker&#x2019;s production that this difference is significant. In contrast, neither vowel duration nor spectral tilt (H1.H2 or H1.A3) is used to differentiate stressed and unstressed syllables by any of the speakers. Finally, both F1 and F2 differentiate stressed and unstressed syllables to a significant extent in the NE speaker&#x2019;s productions, but not in the productions of the EA speaker and JA speaker.</p>
<p>The differences among the three speakers in the observed cues to stress in the experimental stimuli match the generalizations reported for the full set of speakers who participated in the study from which the stimuli were extracted (Almbark <italic>et al.</italic>, <xref ref-type="bibr" rid="cit0001">2014</xref>), which also reports the phonetic realization of stress in L1 EA and JA by the same speakers.</p>
</sec>
<sec id="sec3.2">
<title>3.2. Participants</title>
<p>Participants were recruited by email invitation among the friends and family of graduate students of linguistics from Egypt and Jordan, and among students at the University of York. A total of 42 listeners meeting our inclusion criteria (by native language/dialect, excluding early bilinguals) completed the online perception experiment on a voluntary basis. From these a balanced subset of 36 was selected at random to yield three listener groups by native language: EA, JA or English (NE), with six male and six female listeners in each group. The Arabic-speaking listeners had all studied English for at least 12 years; six had English medium schooling (two EA, four JA); one JA listener was in the UK at the time of taking test.</p>
</sec>
<sec id="sec3.3">
<title>3.3. Procedure</title>
<p>The experiment was run using an online survey tool (SurveyGizmo, <xref ref-type="bibr" rid="cit0029">2019</xref>). Participants first read an information sheet and provided their informed consent to participate; they then completed a questionnaire about age, sex, native language and dialect, and, for L2 listeners, number of years of study of English.</p>
<p>Participants were familiarized with the test paradigm in a training phase; a selection of English stimuli were presented, which differed in stress position as in the main test, using the target word &#x2018;project&#x2019; [&#x02C8;p&#x0279;&#x0252;d&#x0292;&#x025B;kt] ~ [p&#x0279;&#x0259;&#x02C8;d&#x0292;&#x025B;kt]. Participants were asked to answer the following question for each word they heard: &#x201C;Was it PROject (first syllable) or proJECT (second syllable)?&#x201D;. Feedback was given as to whether the provided answer was correct or incorrect.</p>
<p>After the training phase, in the first test phase the 36 sound files produced by the two L2 English speakers (9 target words &#x00D7; 2 stress conditions &#x00D7; 2 speakers = 36) were presented in randomized order. Each sound file was shown on a separate page with the question &#x201C;Is it ___ (first syllable) or ___ (second syllable)?&#x201D; and two answers (e.g., &#x201C;SUBject with stress on the first syllable&#x201D; or &#x201C;subJECT with stress on the second syllable&#x201D;) to choose from, in a binary forced choice. Then, in the second test phase, the 18 sound files produced by the L1 English speaker were presented, following the same procedure as for the first test phase.</p>
<p>We presented all L2 speech in one block, then all L1 speech in a separate block, to restrict the listeners&#x2019; task to word recognition. Randomisation of tokens extracted from L1 and L2 speech in one block might have drawn listeners&#x2019; attention to evaluation of the degree of foreign accent rather than the intelligibility (i.e., recognition) of the utterances as intended.</p>
</sec>
<sec id="sec3.4">
<title>3.4. Analysis</title>
<p>Each response was coded for <italic>accuracy</italic>: responses which matched the intended form of the word as elicited were coded as correct, otherwise as incorrect. Results were explored using binomial generalized linear mixed models (GLMM) using <italic>lme4</italic> (Bates <italic>et al.</italic>, <xref ref-type="bibr" rid="cit0002">2015</xref>) in <italic>R</italic> (Core Team, <xref ref-type="bibr" rid="cit0010">2014</xref>), with <italic>accuracy</italic> as the dependent variable, using likelihood ratio tests to identify the best fit model. The predictions of the model were extracted using <italic>lsmeans</italic> (Lenth, <xref ref-type="bibr" rid="cit0022">2016</xref>) and plots were produced using <italic>ggplot2</italic> (Wickham, <xref ref-type="bibr" rid="cit0033">2009</xref>).</p>
</sec>
</sec>
<sec id="sec4" sec-type="results">
<title>4. RESULTS</title>
<p><xref ref-type="fig" rid="f0003">Figure 3</xref> shows accuracy rates for the three groups of listeners, grouped by stimulus language and elicited position of stress. Accuracy rates are above chance for most participants (where chance would equate to a score of 4 or 5, in a binary forced choice task with a maximum score of 9). Accuracy is above chance for English listeners in response to all stimuli produced by the NE speaker, and there is a ceiling effect for English listeners in response to stimuli elicited with initial stress. Visually, it appears that English listeners are somewhat more accurate than EA listeners, who are in turn somewhat more accurate than JA listeners, but that there is little effect of stimulus language for the Arabic listeners.</p>
<fig id="f0003">
<label>Figure 3</label>
<caption>
<p>Median (bold vertical line within bars) and interquartile range (bar size) of the count of accurate responses for each individual participant, grouped by listener language, stimulus language, and elicited position of stress. EA = Egyptian Arabic; JA = Jordanian Arabic.</p>
</caption>
<graphic xlink:href="loquens_061-g003.tif" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</fig>
<p>However, any variation across listener groups is clearly mediated by variation within listener groups that reflects the elicited position of stress in the word: EA listeners are less accurate at identifying words produced by the English speaker with initial stress; English listeners, in turn, are less accurate at identifying words produced by the Egyptian speaker with final stress. In contrast, accuracy rates of JA listeners show largely overlapping distributions by both position of stress and by stimulus language.</p>
<p>These effects were explored in a series of GLMM models; the best fit model includes fixed factors for stress condition (<italic>stress</italic>), listener language <italic>(listlang)</italic> and stimulus language <italic>(stimlang)</italic>, and all interactions among these three factors, with random intercepts for <italic>participant</italic> and <italic>item</italic>. Separate models were run including the control factors <italic>age</italic>, <italic>sex</italic>, and <italic>device</italic> (encoding participants&#x2019; use of earphones versus external loudspeaker to take the test), but none of these factors improved model fit. The best fit model summary is reported in <xref ref-type="table" rid="t0003">Table 3</xref>. The reference levels for the fixed factors were &#x2018;initial&#x2019; (for <italic>stress</italic>) and &#x2018;EA&#x2019; (for <italic>listlang</italic> and <italic>stimlang</italic>); the model was re-run with &#x2018;JA&#x2019; as reference level to obtain pairwise comparisons (which are reported where relevant in the text).</p>
<table-wrap id="t0003">
<label>Table 3</label>
<caption>
<p>Summary of the best fit GLMM [accuracy ~ listlang * stimlang * stress + (1 | item) + (1 | participant)].</p>
</caption>
<table frame="hsides" rules="groups">
<thead>
<tr>
<th align="left">Fixed effects</th>
<th align="center">Estimate (log odds)</th>
<th align="center"><italic>SE</italic></th>
<th align="center"><italic>z</italic></th>
<th align="center"><italic>p</italic></th>
</tr>
</thead>
<tbody>
<tr>
<td align="left">intercept</td>
<td align="center">1.88724</td>
<td align="center">0.34165</td>
<td align="center">5.524</td>
<td align="center">&#x003C; .000</td>
</tr>
<tr>
<td align="left">stimlangJA</td>
<td align="center">-0.07544</td>
<td align="center">0.38553</td>
<td align="center">-0.196</td>
<td align="center">0.8448</td>
</tr>
<tr>
<td align="left">stimlangNE</td>
<td align="center">-1.30793</td>
<td align="center">0.34590</td>
<td align="center">-3.781</td>
<td align="center">0.0001</td>
</tr>
<tr>
<td align="left">listlangJA</td>
<td align="center">-0.79566</td>
<td align="center">0.41624</td>
<td align="center">-1.912</td>
<td align="center">0.0559</td>
</tr>
<tr>
<td align="left">listlangNE</td>
<td align="center">1.57067</td>
<td align="center">0.61761</td>
<td align="center">2.543</td>
<td align="center">0.0109*</td>
</tr>
<tr>
<td align="left">stresssecond</td>
<td align="center">-0.84735</td>
<td align="center">0.39831</td>
<td align="center">-2.127</td>
<td align="center">0.0333</td>
</tr>
<tr>
<td align="left">stimlangJA:listlangJA</td>
<td align="center">-0.51935</td>
<td align="center">0.49107</td>
<td align="center">-1.058</td>
<td align="center">0.2902</td>
</tr>
<tr>
<td align="left">stimlangNE:listlangJA</td>
<td align="center">0.58733</td>
<td align="center">0.45942</td>
<td align="center">1.278</td>
<td align="center">0.2011</td>
</tr>
<tr>
<td align="left">stimlangJA:listlangNE</td>
<td align="center">-1.03615</td>
<td align="center">0.71293</td>
<td align="center">-1.453</td>
<td align="center">0.1461</td>
</tr>
<tr>
<td align="left">stimlangNE:listlangNE</td>
<td align="center">1.30792</td>
<td align="center">0.79567</td>
<td align="center">1.644</td>
<td align="center">0.1002</td>
</tr>
<tr>
<td align="left">stimlangJA:stresssecond</td>
<td align="center">-0.07017</td>
<td align="center">0.49480</td>
<td align="center">-0.142</td>
<td align="center">0.8872</td>
</tr>
<tr>
<td align="left">stimlangNE:stresssecond</td>
<td align="center">1.57107</td>
<td align="center">0.47366</td>
<td align="center">3.317</td>
<td align="center">0.0009***</td>
</tr>
<tr>
<td align="left">listlangJA:stresssecond</td>
<td align="center">0.25170</td>
<td align="center">0.46696</td>
<td align="center">0.539</td>
<td align="center">0.5898</td>
</tr>
<tr>
<td align="left">listlangNE:stresssecond</td>
<td align="center">-1.40917</td>
<td align="center">0.65985</td>
<td align="center">-2.136</td>
<td align="center">0.0327*</td>
</tr>
<tr>
<td align="left">stimlangJA:listlangJA:stresssecond</td>
<td align="center">1.01822</td>
<td align="center">0.65254</td>
<td align="center">1.560</td>
<td align="center">0.1186</td>
</tr>
<tr>
<td align="left">stimlangNE:listlangJA:stresssecond</td>
<td align="center">-0.85047</td>
<td align="center">0.63254</td>
<td align="center">-1.345</td>
<td align="center">0.1787</td>
</tr>
<tr>
<td align="left">stimlangJA:listlangNE:stresssecond</td>
<td align="center">1.59293</td>
<td align="center">0.84933</td>
<td align="center">1.876</td>
<td align="center">0.0607</td>
</tr>
<tr>
<td align="left">stimlangNE:listlangNE:stresssecond</td>
<td align="center">-1.28632</td>
<td align="center">0.92208</td>
<td align="center">-1.395</td>
<td align="center">0.1630</td>
</tr>
</tbody>
</table>
</table-wrap>
<p>For the dependent variable the reference level in all models was &#x2018;incorrect&#x2019;; the models thus predict the log odds of improved accuracy resulting from a change in <italic>stress</italic> or <italic>stimlang</italic> or <italic>listlang</italic> condition or a combination of these. The predicted marginal means of the model, and 95% confidence intervals around them, are illustrated in <xref ref-type="fig" rid="f0004">Figure 4</xref>; this plot visualizes the significant effects predicted by the model (overlapping confidence intervals indicate an effect which is not significant).</p>
<fig id="f0004">
<label>Figure 4</label>
<caption>
<p>Predicted marginal means (and 95&#x2009;% CI) for the best fit binomial GLMM by listener language, stimulus language (<italic>stimlang</italic>), and position of stress. EA = Egyptian Arabic; JA = Jordanian Arabic; NE = English native speaker.</p>
</caption>
<graphic xlink:href="loquens_061-g004.tif" xmlns:xlink="http://www.w3.org/1999/xlink"/>
</fig>
<p>The best fit model shows no significant three-way interactions and no main effect of stimulus language or stress position. There were no significant interactions between listener language and stimulus language; it is this type of interaction that would indicate an interlanguage intelligibility benefit.</p>
<p>There is a main effect of listener language: English listeners are much more accurate than JA listeners (<italic>z</italic>(1924)&#x2009;= &#x2009;3.967; <italic>p</italic>&#x2009;&#x003C;&#x2009;.000) and also somewhat more accurate than EA listeners (<italic>z</italic>(1924)&#x2009;=&#x2009;2.543; <italic>p</italic>&#x2009;=&#x2009;.010), regardless of speaker language and stress position. This matches the pattern observed for English listeners by Stibbard and Lee (<xref ref-type="bibr" rid="cit0028">2006</xref>).</p>
<p>There is a significant interaction between stress and listener language: English listeners were less accurate at identifying words with final stress across the board, regardless of stimulus language (<italic>z</italic>(1924) = -2.136; <italic>p</italic> = .0327). There was also a significant interaction between stress and stimulus language: words with final stress were less accurately identified by all listeners when produced by the EA speaker, than either the NE speaker (<italic>z</italic>(1924) = 3.317; <italic>p</italic> = .0009) or the JA speaker (<italic>z</italic>(1924) = 2.227; <italic>p</italic> = .0259). We explore these interactions with stress position in the general discussion below.</p>
</sec>
<sec id="sec5" sec-type="discussion">
<title>5. DISCUSSION</title>
<p>Our specific research question was to explore a possible interlanguage intelligibility benefit in perception of English word stress; that is, to test the hypothesis that L2 listeners will more accurately interpret English word stress when produced by other L2 speakers. We found no evidence to support this hypothesis in this study, as there are no significant interactions between any levels of listener language and stimulus language in our data. This also rules out the type of interlanguage intelligibility benefit found by Stibbard and Lee (<xref ref-type="bibr" rid="cit0028">2006</xref>), where L2 listeners perform better when listening to speakers from the same L2 background: the distribution of accuracy rates for EA listeners in response to EA stimuli overlaps with that observed in response to JA stimuli (and likewise, the distribution of accuracy rates for JA listeners in response to JA stimuli overlaps with that in response to EA stimuli). We thus find no evidence for an interlanguage intelligibility benefit based on phonetic realization of stress.</p>
<p>Our results replicate the finding of Stibbard and Lee (<xref ref-type="bibr" rid="cit0028">2006</xref>) who also found that English listeners performed better in a sentence recognition task across the board, in comparison to L2 listeners. Our study extends this finding to include recognition of lexical items differentiated solely by stress, in response to stimuli which bear cues to stress only, without any additional enhancement in cues due to phrase-level accent. We attribute this finding to the ability of L1 listeners to make use of &#x2018;top-down&#x2019; structural and/or lexico-semantic cues in perception of stress; in the present study this could be because the native English listeners are more familiar with the lexical items used as stimuli than L2 learners are. The lower accuracy of the L2 listeners in our results, across the board, mirrors the findings of other studies which showed that L2 listeners are less reliant on &#x2018;top-down&#x2019; cues; in the present study this may be a direct effect of reduced familiarity with some of the lexical items, and/or reduced of awareness of the existence of stress near-minimal pairs in English. These competing explanations could be explored in future research by using pseudoword stimuli or by controlling for L2 learners&#x2019; vocabulary size.</p>
<p>The study shows two significant interactions of listener/speaker language with the position of stress. The first of these is that NE listeners displayed lower accuracy in response to words with final stress, regardless of speaker language. The expected NE realization in these words has vowel reduction in the first syllable, to schwa [&#x0259;] in 7 out of 9 of our stimuli, and to [&#x026A;] in the other two cases (see <xref ref-type="table" rid="t0001">Table 1</xref>). Vowel reduction is in fact the primary cue to stress for English listeners (Cutler, <xref ref-type="bibr" rid="cit0012">1986</xref>), so this result suggests that the reduced vowel reduction in the stimuli produced by the two L2 speakers (illustrated in <xref ref-type="fig" rid="f0002">Figure 2</xref>) may indeed have contributed to lower intelligibility of their productions by the NE listeners. The second significant interaction with stress position was that all listeners were less accurate in their interpretation of the EA speaker&#x2019;s productions of words with final stress. We attribute this reduced accuracy to the reduced differentiation of stressed and unstressed syllables in the productions of the EA speaker (see <xref ref-type="fig" rid="f0001">Figure 1</xref>); this lack of differentiation may in turn result from the previously reported conflation of word- and phrase-level stress in this dialect (Hellmuth, <xref ref-type="bibr" rid="cit0019">2007</xref>). Taken together we interpret these interactions as evidence that non-target-like phonetic realization of stress can result in lower intelligibility of L2 speakers&#x2019; productions for both L1 and L2 listeners in certain contexts.</p>
</sec>
<sec id="sec6" sec-type="conclusions">
<title>6. CONCLUSION</title>
<p>The aim of this paper was to explore a possible interlanguage intelligibility benefit for L2 listeners in perception of stress, due to potential transfer of L1 patterns of phonetic realization of stress into L2 productions. The results did not show any interlanguage intelligibility benefit but did confirm the previous finding of an overall advantage for L1 English listeners in lexical recognition tasks, which we attribute to the L1 listeners&#x2019; ability to make use of top-down lexical knowledge in perception of stress. This strategy supports accurate recognition in the face of non-target-like phonetic cues to stress encountered in L2 English productions, but we show that non-target-like cues can result in reduced intelligibility of L2 speakers when the primary cue expected by L1 listeners (here, vowel reduction) is the same cue that the L2 speaker fails to produce to a target-like extent.</p>
</sec>
</body>
<back>
<ref-list>
<title>REFERENCES</title>
<ref id="cit0001">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Almbark</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>Bouchhioua</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Hellmuth</surname>
<given-names>S.</given-names>
</name>
</person-group>
<year>2014</year>
<chapter-title>Acquiring the phonetics and phonology of English word stress: Comparing learners from different L1 backgrounds</chapter-title>
<source>Proceedings of the International Symposium on the Acquisition of Second Language Speech, Concordia Working Papers in Applied Linguistics</source>
<volume>5</volume>
</mixed-citation>
</ref>
<ref id="cit0002">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bates</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Maechler</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Bolker</surname>
<given-names>B.</given-names>
</name>
<name>
<surname>Walker</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Fitting linear mixed-effects models using lme4</article-title>
<source>Journal of Statistical Software</source>
<year>2015</year>
<volume>67</volume>
<issue>1</issue>
<fpage>1</fpage>
<lpage>48</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.18637/jss.v067.i01">https://doi.org/10.18637/jss.v067.i01</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0003">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Beckman</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Edwards</surname>
<given-names>J.</given-names>
</name>
</person-group>
<year>1994</year>
<chapter-title>Articulatory evidence for differentiating stress categories</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Keating</surname>
<given-names>P.</given-names>
</name>
</person-group>
<source>Phonological structure and phonetic form: Papers in Laboratory Phonology III</source>
<fpage>7</fpage>
<lpage>33</lpage>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1017/CBO9780511659461.002">https://doi.org/10.1017/CBO9780511659461.002</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0004">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bent</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Bradlow</surname>
<given-names>A. R.</given-names>
</name>
</person-group>
<article-title>The interlanguage speech intelligibility benefit</article-title>
<source>The Journal of the Acoustical Society of America</source>
<year>2003</year>
<volume>114</volume>
<issue>3</issue>
<fpage>1600</fpage>
<lpage>1610</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1603234">https://doi.org/10.1121/1.1603234</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0005">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Bouchhioua</surname>
<given-names>N.</given-names>
</name>
</person-group>
<year>2008</year>
<source>The acoustic correlates of stress and accent in Tunisian Arabic: A comparative study with English</source>
<comment>[Unpublished PhD dissertation]</comment>
<publisher-name>University of Carthage</publisher-name>
<publisher-loc>Tunisia</publisher-loc>
</mixed-citation>
</ref>
<ref id="cit0006">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Bouchhioua</surname>
<given-names>N.</given-names>
</name>
</person-group>
<article-title>Typological variation in the phonetic realization of lexical and phrasal stress: Southern British English vs. Tunisian Arabic</article-title>
<source>Loquens</source>
<year>2016</year>
<volume>3</volume>
<issue>2</issue>
<fpage>e034.</fpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3989/loquens.2016.034">https://doi.org/10.3989/loquens.2016.034</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0007">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Chahal</surname>
<given-names>D.</given-names>
</name>
<name>
<surname>Hellmuth</surname>
<given-names>S.</given-names>
</name>
</person-group>
<year>2015</year>
<chapter-title>Comparing the intonational phonology of Lebanese and Egyptian Arabic</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<source>Prosodic typology</source>
<volume>2</volume>
<fpage>365</fpage>
<lpage>404</lpage>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1093/acprof:oso/9780199567300.003.0013">https://doi.org/10.1093/acprof:oso/9780199567300.003.0013</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0008">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Chrabaszcz</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Winn</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Lin</surname>
<given-names>C. Y.</given-names>
</name>
<name>
<surname>Idsardi</surname>
<given-names>W. J.</given-names>
</name>
</person-group>
<article-title>Acoustic cues to perception of word stress by English, Mandarin, and Russian speakers</article-title>
<source>Journal of Speech, Language, and Hearing Research</source>
<year>2014</year>
<volume>57</volume>
<issue>4</issue>
<fpage>1468</fpage>
<lpage>1479</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1044/2014_JSLHR-L-13-0279">https://doi.org/10.1044/2014_JSLHR-L-13-0279</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0009">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cole</surname>
<given-names>J.</given-names>
</name>
<name>
<surname>Mo</surname>
<given-names>Y.</given-names>
</name>
<name>
<surname>Hasegawa-Johnson</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Signal-based and expectation-based factors in the perception of prosodic prominence</article-title>
<source>Laboratory Phonology</source>
<year>2010</year>
<volume>1</volume>
<issue>2</issue>
<fpage>425</fpage>
<lpage>452</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1515/labphon.2010.022">https://doi.org/10.1515/labphon.2010.022</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0010">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Core Team</surname>
<given-names>R.</given-names>
</name>
</person-group>
<year>2014</year>
<source>R: A language and environment for statistical computing</source>
<publisher-loc>Vienna</publisher-loc>
<publisher-name>Austria</publisher-name>
<comment>Retrieved from <ext-link ext-link-type="uri" xlink:href="http://www.R-project.org/">http://www.R-project.org/</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0011">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Cruttenden</surname>
<given-names>A.</given-names>
</name>
</person-group>
<year>2006</year>
<chapter-title>The de-accenting of old information: A cognitive universal</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Bernini</surname>
<given-names>G.</given-names>
</name>
<name>
<surname>Schwartz</surname>
<given-names>M. L.</given-names>
</name>
</person-group>
<source>Pragmatic organization of discourse in the languages of Europe</source>
<fpage>311</fpage>
<lpage>355</lpage>
<publisher-loc>Berlin</publisher-loc>
<publisher-name>Mouton de Gruyter</publisher-name>
</mixed-citation>
</ref>
<ref id="cit0012">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Cutler</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Forebear is a homophone: Lexical prosody does not constrain lexical access</article-title>
<source>Language and Speech</source>
<year>1986</year>
<volume>29</volume>
<issue>3</issue>
<fpage>201</fpage>
<lpage>220</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/002383098602900302">https://doi.org/10.1177/002383098602900302</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0013">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>de Jong</surname>
<given-names>K.</given-names>
</name>
<name>
<surname>Zawaydeh</surname>
<given-names>B. A.</given-names>
</name>
</person-group>
<article-title>Stress, duration, and intonation in Arabic word-level prosody</article-title>
<source>Journal of Phonetics</source>
<year>1999</year>
<volume>27</volume>
<issue>1</issue>
<fpage>3</fpage>
<lpage>22</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1006/jpho.1998.0088">https://doi.org/10.1006/jpho.1998.0088</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0014">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dupoux</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Pallier</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Sebasti&#x00E1;n-Gall&#x00E9;s</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Mehler</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>A destressing &#x2018;deafness&#x2019; in French?</article-title>
<source>Journal of Memory and Language</source>
<year>1997</year>
<volume>36</volume>
<fpage>406</fpage>
<lpage>421</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1006/jmla.1996.2500">https://doi.org/10.1006/jmla.1996.2500</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0015">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Dupoux</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Sebasti&#x00E1;n-Gall&#x00E9;s</surname>
<given-names>N.</given-names>
</name>
<name>
<surname>Navarrete</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Peperkamp</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>Persistent stress &#x201C;deafness&#x201D;: The case of French learners of Spanish</article-title>
<source>Cognition: International Journal of Cognitive Science</source>
<year>2008</year>
<volume>106</volume>
<fpage>682</fpage>
<lpage>706</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/j.cognition.2007.04.001">https://doi.org/10.1016/j.cognition.2007.04.001</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0016">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Eriksson</surname>
<given-names>A.</given-names>
</name>
<name>
<surname>Thunberg</surname>
<given-names>G. C.</given-names>
</name>
<name>
<surname>Traunm&#x00FC;ller</surname>
<given-names>H.</given-names>
</name>
</person-group>
<chapter-title>Syllable prominence: A matter of vocal effort, phonetic distinctness and top-down processing</chapter-title>
<source>EUROSPEECH-2001</source>
<year>2001</year>
<fpage>399</fpage>
<lpage>402</lpage>
</mixed-citation>
</ref>
<ref id="cit0017">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Fry</surname>
<given-names>D.</given-names>
</name>
</person-group>
<article-title>Duration and intensity as physical correlates of linguistic stress</article-title>
<source>Journal of the Acoustical Society of America</source>
<year>1955</year>
<volume>27</volume>
<issue>4</issue>
<fpage>765</fpage>
<lpage>768</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1908022">https://doi.org/10.1121/1.1908022</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0018">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Gordon</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Roettger</surname>
<given-names>T.</given-names>
</name>
</person-group>
<article-title>Acoustic correlates of word stress: A cross-linguistic survey</article-title>
<source>Linguistics Vanguard</source>
<year>2017</year>
<volume>3</volume>
<issue>1</issue>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1515/lingvan-2017-0007">https://doi.org/10.1515/lingvan-2017-0007</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0019">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Hellmuth</surname>
<given-names>S.</given-names>
</name>
</person-group>
<article-title>The relationship between prosodic structure and pitch accent distribution: Evidence from Egyptian Arabic</article-title>
<source>The Linguistic Review</source>
<year>2007</year>
<volume>24</volume>
<issue>2&#x2013;3</issue>
<fpage>289</fpage>
<lpage>314</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1515/TLR.2007.011">https://doi.org/10.1515/TLR.2007.011</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0020">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Jun</surname>
<given-names>S.-A</given-names>
</name>
</person-group>
<year>2014</year>
<chapter-title>Prosodic typology: By prominence type, word prosody, and macro-rhythm</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Jun</surname>
<given-names>S.-A.</given-names>
</name>
</person-group>
<source>Prosodic typology Volume II</source>
<fpage>520</fpage>
<lpage>539</lpage>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Oxford University Press</publisher-name>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1093/acprof:oso/9780199567300.001.0001">https://doi.org/10.1093/acprof:oso/9780199567300.001.0001</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0021">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ladd</surname>
<given-names>D. R.</given-names>
</name>
</person-group>
<year>2008</year>
<source>Intonational phonology</source>
<edition>2</edition>
<publisher-loc>Cambridge</publisher-loc>
<publisher-name>Cambridge University Press</publisher-name>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1017/CBO9780511808814">https://doi.org/10.1017/CBO9780511808814</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0022">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Lenth</surname>
<given-names>R. V.</given-names>
</name>
</person-group>
<article-title>Least-squares means: The R package lsmeans</article-title>
<source>Journal of Statistical Software</source>
<year>2016</year>
<volume>69</volume>
<issue>1</issue>
<fpage>1</fpage>
<lpage>33</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.18637/jss.v069.i01">https://doi.org/10.18637/jss.v069.i01</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0023">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Mattys</surname>
<given-names>S.</given-names>
</name>
<name>
<surname>White</surname>
<given-names>L.</given-names>
</name>
<name>
<surname>Melhorn</surname>
<given-names>J.</given-names>
</name>
</person-group>
<article-title>Integration of multiple segentation cues: A hierarchical framework</article-title>
<source>Journal of Experimental Psychology: General</source>
<year>2005</year>
<volume>134</volume>
<fpage>477</fpage>
<lpage>500</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/0096-3445.134.4.477">https://doi.org/10.1037/0096-3445.134.4.477</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0024">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Ohala</surname>
<given-names>J.</given-names>
</name>
</person-group>
<year>1993</year>
<chapter-title>The phonetics of sound change</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Jones</surname>
<given-names>C.</given-names>
</name>
</person-group>
<source>Historical linguistics: Problems and perspectives</source>
<fpage>237</fpage>
<lpage>278</lpage>
<publisher-loc>London</publisher-loc>
<publisher-name>Longman</publisher-name>
</mixed-citation>
</ref>
<ref id="cit0025">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Qin</surname>
<given-names>Z.</given-names>
</name>
<name>
<surname>Chien</surname>
<given-names>Y.-F.</given-names>
</name>
<name>
<surname>Tremblay</surname>
<given-names>A.</given-names>
</name>
</person-group>
<article-title>Processing of word-level stress by Mandarin-speaking second language learners of English</article-title>
<source>Applied Psycholinguistics</source>
<year>2017</year>
<volume>38</volume>
<issue>3</issue>
<fpage>541</fpage>
<lpage>570</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1017/S0142716416000321">https://doi.org/10.1017/S0142716416000321</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0026">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Roettger</surname>
<given-names>T.</given-names>
</name>
<name>
<surname>Gordon</surname>
<given-names>M.</given-names>
</name>
</person-group>
<article-title>Methodological issues in the study of word stress correlates</article-title>
<source>Linguistics Vanguard</source>
<year>2017</year>
<volume>3</volume>
<issue>1</issue>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1515/lingvan-2017-0006">https://doi.org/10.1515/lingvan-2017-0006</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0027">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Selinker</surname>
<given-names>L.</given-names>
</name>
</person-group>
<year>1972</year>
<chapter-title>Interlanguage</chapter-title>
<source>IRAL-International Review of Applied Linguistics in Language Teaching</source>
<volume>10</volume>
<issue>1&#x2013;4</issue>
<fpage>209</fpage>
<lpage>232</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1515/iral.1972.10.1-4.209">https://doi.org/10.1515/iral.1972.10.1-4.209</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0028">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Stibbard</surname>
<given-names>R. M.</given-names>
</name>
<name>
<surname>Lee</surname>
<given-names>J.-I.</given-names>
</name>
</person-group>
<article-title>Evidence against the mismatched interlanguage speech intelligibility benefit hypothesis</article-title>
<source>The Journal of the Acoustical Society of America</source>
<year>2006</year>
<volume>120</volume>
<issue>1</issue>
<fpage>433</fpage>
<lpage>442</lpage>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.2203595">https://doi.org/10.1121/1.2203595</ext-link>
</comment>
</nlm-citation>
</ref>
<ref id="cit0029">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<collab>SurveyGizmo</collab>
</person-group>
<year>2019</year>
<source>SurveyGizmo | Enterprise Online Survey Software &#x0026; Tools</source>
<comment>Retrieved from <ext-link ext-link-type="uri" xlink:href="https://www.surveygizmo.com/">https://www.surveygizmo.com/</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0030">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>van Heuven</surname>
<given-names>V.</given-names>
</name>
<name>
<surname>Sluijter</surname>
<given-names>A.</given-names>
</name>
</person-group>
<year>1996</year>
<chapter-title>Notes on the phonetics of word prosody</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Goedmans</surname>
<given-names>R.</given-names>
</name>
<name>
<surname>van der Hulst</surname>
<given-names>H.</given-names>
</name>
<name>
<surname>Visch</surname>
<given-names>E.</given-names>
</name>
</person-group>
<source>Stress patterns of the world. Part 1: Background</source>
<fpage>233</fpage>
<lpage>269</lpage>
<publisher-loc>The Hague</publisher-loc>
<publisher-name>Holland Academic Graphics</publisher-name>
</mixed-citation>
</ref>
<ref id="cit0031">
<nlm-citation publication-type="journal">
<person-group person-group-type="author">
<name>
<surname>Wagner</surname>
<given-names>P.</given-names>
</name>
</person-group>
<article-title>Great expectations-introspective vs. perceptual prominence ratings and their acoustic correlates</article-title>
<source>INTERSPEECH-2005</source>
<year>2005</year>
<fpage>2381</fpage>
<lpage>2384</lpage>
</nlm-citation>
</ref>
<ref id="cit0032">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Watson</surname>
<given-names>J. C. E.</given-names>
</name>
</person-group>
<year>2011</year>
<chapter-title>Word stress in Arabic</chapter-title>
<person-group person-group-type="editor">
<name>
<surname>Oostendorp</surname>
<given-names>M.</given-names>
</name>
<name>
<surname>Ewen</surname>
<given-names>C.</given-names>
</name>
<name>
<surname>Hume</surname>
<given-names>E.</given-names>
</name>
<name>
<surname>Rice</surname>
<given-names>K.</given-names>
</name>
</person-group>
<source>The Blackwell Companion to Phonology</source>
<volume>5</volume>
<fpage>2990</fpage>
<lpage>3018</lpage>
<publisher-loc>Oxford</publisher-loc>
<publisher-name>Blackwell</publisher-name>
<comment>
<ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1002/9781444335262.wbctp0124">https://doi.org/10.1002/9781444335262.wbctp0124</ext-link>
</comment>
</mixed-citation>
</ref>
<ref id="cit0033">
<mixed-citation publication-type="book">
<person-group person-group-type="author">
<name>
<surname>Wickham</surname>
<given-names>H.</given-names>
</name>
</person-group>
<year>2009</year>
<source>ggplot2: Elegant graphics for data analysis</source>
<publisher-name>Springer Science &#x0026; Business Media</publisher-name>
</mixed-citation>
</ref>
</ref-list>
</back>
</article>
