<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD Journal Publishing DTD v3.0 20080202//EN" "journalpublishing3.dtd">
<article article-type="research-article" dtd-version="3.0" xml:lang="en" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">
	<front>
		<journal-meta>
			<journal-id journal-id-type="publisher-id">Loquens</journal-id>
			<journal-title-group>
				<journal-title>Loquens</journal-title>
				<abbrev-journal-title>Loquens</abbrev-journal-title>
			</journal-title-group>
			<issn pub-type="epub">2386-2637</issn>
			<publisher>
				<publisher-name>Consejo Superior de Investigaciones Científicas</publisher-name>
			</publisher>
		</journal-meta>
		<article-meta>
			<article-id pub-id-type="publisher-id">loquens.2017.044</article-id>
			<article-id pub-id-type="doi">10.3989/loquens.2017.044</article-id>
			<article-categories>
				<subj-group subj-group-type="heading">
					<subject>Articles</subject>
				</subj-group>
			</article-categories>
			<title-group>
				<article-title>Individual variability in cue weighting for first-language vowels</article-title>
				<trans-title-group xml:lang="es">
					<trans-title>Variabilidad individual en el peso de claves en vocales de una primera lengua</trans-title>
				</trans-title-group>
				<alt-title alt-title-type="running-head">Individual variability in cue weighting for first-language vowels</alt-title>
			</title-group>
			<contrib-group>
				<contrib contrib-type="author" corresp="yes">
					<name>
						<surname>Ghaffarvand Mokari</surname>
						<given-names>Payam</given-names>
					</name>
					<xref ref-type="corresp" rid="cor1"/>
				<xref ref-type="aff" rid="aff1"/>
				</contrib>
			<contrib contrib-type="author" corresp="yes">
					<name>
						<surname>Werner </surname>
						<given-names>Stefan</given-names>
					</name>
					<xref ref-type="corresp" rid="cor2"/>
				<xref ref-type="aff" rid="aff1"/>
				</contrib>
				<aff id="aff1">University of Eastern Finland </aff>
			</contrib-group>
			<author-notes>
				<corresp id="cor1">e-mail: <email xlink:href="payam.ghaffarvand@uef.fi">payam.ghaffarvand@uef.fi</email> ORCID: <ext-link ext-link-type="uri" xlink:href="https://orcid.org/0000-0002-1816-2783">https://orcid.org/0000-0002-1816-2783</ext-link></corresp>
				<corresp id="cor2">e-mail: <email xlink:href="stefan.werner@uef.fi">stefan.werner@uef.fi</email> ORCID: <ext-link ext-link-type="uri" xlink:href="https://orcid.org/0000-0001-5176-8114">https://orcid.org/0000-0001-5176-8114</ext-link></corresp>
			</author-notes>
			<pub-date pub-type="epub">
				<day>31</day>
				<month>07</month>
				<year>2017</year>
			</pub-date>
			<pub-date pub-type="collection">
				<year>2017</year>
			</pub-date>
			<volume>4</volume>
			<issue>2</issue>
			<elocation-id content-type="doi">10.3989/loquens.2017.044</elocation-id>
			<history>
				<date date-type="received">
					<day>21</day>
					<month>01</month>
					<year>2017</year>
				</date>
				<date date-type="accepted">
					<day>07</day>
					<month>07</month>
					<year>2017</year>
				</date>
				<date date-type="published online">
					<day>21</day>
					<month>02</month>
					<year>2018</year>
				</date>
			</history>
			<permissions>
				<copyright-statement>© 2017 CSIC</copyright-statement>
				<copyright-year>2017</copyright-year>
				<license license-type="open-access" xlink:href="http://creativecommons.org/licenses/by/3.0/es/deed.es">
					<license-p>This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY) Spain 3.0</license-p>
				</license>
			</permissions>
			<abstract xml:lang="en" id="abstract01">
				<title>ABSTRACT</title>
				<p>This study investigates the use of different cues in discrimination of Azerbaijani /œ/ and /ɯ/ vowels. Regarding the large overlap in <italic>f</italic>
		<sub>1</sub>–<italic>f</italic>
		<sub>2</sub> vowel space in the production of these vowels, this study researched for other possible cues in their categorization. Twenty native Azerbaijani listeners were tested through a perceptual identification test. Since <italic>f</italic>
		<sub>2</sub> was weighted more consistently through the experiment, it is suggested that <italic>f</italic>
		<sub>2</sub> is the primary cue in discriminating of this vowel pair. We observed individual differences in the perceptual weighting of <italic>f</italic>
		<sub>2</sub> and <italic>f</italic>
		<sub>3</sub> among the listeners. Although most of the participants gave more weight to <italic>f</italic>
		<sub>2</sub>, some others weighted <italic>f</italic>
		<sub>3</sub> heavier than <italic>f</italic>
		<sub>2</sub> or gave weight to both cues equally. These findings expand the knowledge on perceptual cue weighting and point the importance of examining cue weighting at the individual level.</p>
			</abstract>
			<trans-abstract xml:lang="es" id="abstract02">
				<title>RESUMEN</title>
				<p><italic>Variabilidad individual en el peso de claves en vocales de una primera lengua.—</italic>Este estudio investiga el uso de diferentes claves en la discriminación de las vocales /œ/ y /ɯ/ del azerbaiyano. Teniendo en cuenta el considerable solapamiento que en la producción de estas vocales se produce en el espacio vocálico de <italic>f</italic>
		<sub>1</sub>–<italic>f</italic>
		<sub>2</sub>, en este estudio se han investigado otras posibles claves en su categorización. Veinte oyentes, hablantes nativos de azerbaiyano, participaron en un test perceptivo de identificación. Puesto que se otorgó a <italic>f</italic>
		<sub>2</sub> un peso más consistente a lo largo del experimento, se sugiere que <italic>f</italic>
		<sub>2</sub> es la clave primaria para identificar este par vocálico, si bien observamos diferencias individuales entre los hablantes en el peso atribuido a <italic>f</italic>
		<sub>2</sub> y <italic>f</italic>
		<sub>3</sub>. Aunque la mayor parte de los participantes otorgaron más peso a <italic>f</italic>
		<sub>2</sub>, otros se lo atribuyeron a <italic>f</italic>
		<sub>3</sub> en mayor medida que a <italic>f</italic>
		<sub>2</sub>, o se apoyaron en las dos claves por igual. Estos resultados amplían el conocimiento que se posee sobre el peso de las distintas clases perceptivas y ponen de manifiesto la importancia que tiene examinar el peso de tales claves en cada individuo. </p>
			</trans-abstract>
			<kwd-group xml:lang="en">
				<title>KEYWORDS</title>
				<kwd>individual differences</kwd>
				<kwd>cue weighting</kwd>
				<kwd>perception</kwd>
				<kwd>Azerbaijani vowels</kwd>
			</kwd-group>
			<kwd-group xml:lang="es">
				<title>PALABRAS CLAVE</title>
				<kwd>diferencias individuales</kwd>
				<kwd>peso de las claves</kwd>
				<kwd>percepción</kwd>
				<kwd>vocales del azerbaiyano</kwd>
			</kwd-group>
		</article-meta>
	</front>
	<body>
		<sec id="S1">
		<label>1.</label>
			<title>INTRODUCTION</title>
			<p>There are multiple acoustic dimensions that define speech categories. During spoken language comprehension, listeners categorize speech sounds based on these continuous acoustic cues. Listeners need to determine which cues are relevant to pay attention to, and what relative importance each cue has in order to assign more weight to that cue. Several studies have attempted to find the acoustic dimensions that are important in the discrimination of different speech sounds. <xref ref-type="bibr" rid="b33">Morrison (2013)</xref> provides a review of theories related to dynamic aspects of vowel perception. <xref ref-type="bibr" rid="b42">Strange and Jenkins (2013)</xref>, in their Dynamic Specification model, mention that the most important cues to vowel identity are in the spectro-temporal patterns of consonant–vowel and vowel–consonant formant transitions.</p>
		<p>Among the early studies on the role of different cues in the perception of vowels is the study by <xref ref-type="bibr" rid="b5">Bennett (1968)</xref>, which investigated the relative importance of the spectral and temporal cues in the discrimination of pairs of English and German vowels. He suggested that the importance of the temporal cue is inversely proportional to the distance between the qualities of a given pair of vowels. His results showed that spectral form is, in general, more important than duration in vowel recognition in both English and German, and it is only when two vowels are very close in quality that the duration cue is more important for their discrimination. <xref ref-type="bibr" rid="b3">Ainsworth (1972)</xref> used sets of synthetic vowel sounds that differed in first-formant frequency, second-formant frequency, and duration, and investigated the effect of these cues on the identification of vowels. He found that listeners’ judgments depended on all of 
 these factors; however, duration was a relatively more important cue for vowels located in the centre of the <italic>f</italic>
			<sub>1</sub>–<italic>f</italic>
			<sub>2</sub> space where a vowel might more readily be confused with one of its neighbours.</p>
		<p>According to <xref ref-type="bibr" rid="b25">Idemaru et al. (2012)</xref>, “whereas any of the acoustic dimensions may play a role in phonetic categorization, they are not necessarily perceptually equivalent”. Giving greater perceptual weight to some of the acoustic dimensions is referred to as <italic>cue weighting</italic> (<xref ref-type="bibr" rid="b24">Holt &amp; Lotto, 2006</xref>; <xref ref-type="bibr" rid="b17">Francis, Kaganovich, &amp; Driscoll-Huber, 2008</xref>; <xref ref-type="bibr" rid="b25">Idemaru et al., 2012</xref>). <xref ref-type="bibr" rid="b23">Hillenbrand, Clark, and Houde (2000)</xref> found that English listeners give more weight to the spectral than the temporal dimension in categorizing English [i] and [ɪ] vowels. It has also been found that in discrimination of voiced and voiceless bilabial stop consonants at the syllable initial position, voice onset time (VOT) is more strongly weighted by English listeners and fundamental frequency (<italic>f</
 italic>
			<sub>0</sub>) of the following vowel is used as a secondary cue in the discrimination of this pair (<xref ref-type="bibr" rid="b1">Abramson &amp; Lisker, 1985</xref>; <xref ref-type="bibr" rid="b17">Francis et al., 2008</xref>). <xref ref-type="bibr" rid="b24">Holt and Lotto (2006)</xref> suggest that dimensions that are highly related to category identity need to be more strongly perceptually weighted than those less predictive of category identity. These acoustic dimensions are sometimes weighted differently among listeners.</p>
		<p>There is not extensive research on individual differences in cue-weighting strategies in speech perception (see, e.g., <xref ref-type="bibr" rid="b4">Allen, Miller, &amp; DeSteno, 2003</xref>; <xref ref-type="bibr" rid="b21">Haggard, Ambler, &amp; Callow, 1970</xref>; <xref ref-type="bibr" rid="b22">Hazan &amp; Rosen, 1991</xref>; <xref ref-type="bibr" rid="b25">Idemaru, Holt, &amp; Seltman, 2012</xref>; <xref ref-type="bibr" rid="b29">Kong &amp; Edwards, 2011</xref>, <xref ref-type="bibr" rid="b30">2016</xref>; <xref ref-type="bibr" rid="b37">Raizada, Tsao, Liu, &amp; Kuhl, 2010</xref>; <xref ref-type="bibr" rid="b39">Shultz, Francis, &amp; Llanos, 2012</xref>). Some studies have shown that individual variability exists in cue-weighting strategies (<xref ref-type="bibr" rid="b21">Haggard et al., 1970</xref>; <xref ref-type="bibr" rid="b22">Hazan &amp; Rosen, 1991</xref>; <xref ref-type="bibr" rid="b25">Idemaru et al., 2012</xref>). For instance, <xref ref-type="bibr" rid="b39">
 Shultz et al. (2012)</xref> reported individual differences in the extent of reliance on the secondary cue (<italic>f</italic>
			<sub>0</sub>) in discrimination of /b/ and /p/. <xref ref-type="bibr" rid="b25">Idemaru et al. (2012)</xref> suggest that examining individual differences in perceptual cue weighting in a situation where different dimensions provide similar informativeness provides an opportunity to better understand listeners’ sensitivity to distributional characteristics of acoustic dimensions in categorization of speech sounds. For instance, <xref ref-type="bibr" rid="b8">Chládková, Hamann, Williams, and Hellmuth (2016)</xref> found that <italic>f</italic>
			<sub>2</sub> slope direction is used as a cue (additional to midpoint formant values) to distinguish /iː/ from /uː/ by British English listeners.</p>
			<sec id="S1.1">
				<label>1.1.</label>
				<title>Vowel perception</title>
				<p>Through the history of speech research, formants have played an important role in the studies of vowel perception and acoustic descriptions of vowels. <xref ref-type="bibr" rid="b13">Fant (1960)</xref> indicated the importance of formant frequencies as the prime determinants of the spectral envelope of oral vowels, suggesting that the complex spectra of vowel-like sounds could be uniquely indexed with relatively few parameters. Since formant amplitude appeared to be redundant with formant frequency (<xref ref-type="bibr" rid="b12">Fant, 1956</xref>; <xref ref-type="bibr" rid="b40">Stevens, 1998</xref>) and because formant bandwidth appeared to have little influence on perception (<xref ref-type="bibr" rid="b28">Klatt, 1982</xref>), the focus in speech perception studies was placed on formant frequencies as correlates of perceptual vowel identification.</p>
		<p>Some other studies have shown that general spectral shape is a good correlate to the measures of psychoacoustic distance between vowel-like stimuli (<xref ref-type="bibr" rid="b6">Bladon &amp; Lindblom, 1981</xref>; <xref ref-type="bibr" rid="b36">Pols, van der Kamp, &amp; Plomp, 1969</xref>). In this approach, it is suggested that listeners compare vowel spectra to find the closest match to an internal representation of the corresponding vowel categories. Therefore, formant peaks are not treated differently from other spectral properties and all spectral components are given weight. However, <xref ref-type="bibr" rid="b27">Kiefte and Kluender (2008)</xref> proposed that listeners ignore spectral shape properties in the identification of synthetic monophthongs when the target stimuli were embedded in a sentence.</p>
		<p>In addition to formant-based and spectral shape approaches another approach in vowel perception research is the concept of spectral features based on an intermediate representation. Some studies mention that auditory <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> are perceptually interrelated. <xref ref-type="bibr" rid="b10">Delattre, Liberman, Cooper, and Gerstman (1952)</xref> noted that it was possible to produce acceptable versions of French vowels with the Pattern Playback, with two energy bands close to measured <italic>f</italic>
			<sub>1</sub> and <italic>f</italic>
			<sub>2</sub> of naturally produced vowels. This higher <italic>f</italic>
			<sub>2</sub> has been called the <italic>effective </italic>f2 or f2<italic> prime</italic> (<xref ref-type="bibr" rid="b14">Fant, 1973</xref>; <xref ref-type="bibr" rid="b15">Fant &amp; Risberg, 1963)</xref>. <xref ref-type="bibr" rid="b7">Chistovich and Lublinskaya (1979)</xref> proposed that formant peaks closer than 3.0–3.5 Bark are merged into a single perceived spectral prominence.</p>
		<p><xref ref-type="bibr" rid="b18">Fujimura (1967</xref>) studied the perception of high vowels in Swedish to investigate the theories of formant integration into <italic>f</italic>
			<sub>2</sub> prime. The notion of wideband integration of <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> was criticized based on his results and he instead proposed that both <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> make independent contributions even when they are separated by less than 3.0 Bark. <xref ref-type="bibr" rid="b38">Rosner and Pickering (1994, pp. 151–152)</xref> give experimental evidence indicating that it is not likely that higher formants merge auditorily into a single effective perceptual feature. <xref ref-type="bibr" rid="b35">Nearey and Kiefte (2003)</xref> used a neural network to model spectral integration similar to that proposed by <italic>f</italic>
			<sub>2</sub> prime models in order to reduce a large three-dimensional vowel continuum to two effective formants or parameters; however, this attempt was not successful. A three-dimensional formant-based representation performed substantially better in predicting listeners’ vowel judgments than any two-dimensional representation that could be discovered with the neural network. This again supports <xref ref-type="bibr" rid="b18">Fujimura’s (1967)</xref> hypothesis that vowel perception cannot be explained with two parameters alone.</p>
			</sec>
			<sec id="S1.2">
				<label>1.2.</label>
				<title>Azerbaijani vowels</title>
				<p>The Azerbaijani belongs to the western group of the southwestern, or Oghuz, branch of the Turkic language family and is mainly spoken in Azerbaijan and Iran. Among nonPersian languages in Iran, Azerbaijani, with approximately 15–20 million native speakers, has the largest number of speakers (<xref ref-type="bibr" rid="b9">Crystal, 2010</xref>). Azerbaijani has nine vowels, /æ ɑ o e œ ɯ u i y/, with no length distinction (<xref ref-type="fig" rid="F1">Figure 1</xref>).</p>
				<fig id="F1">
					<label>Figure 1.</label>
					<caption>
						<title>A vowel chart of Azerbaijani (<xref ref-type="bibr" rid="b20">Ghaffarvand Mokari &amp; Werner, 2017</xref>).</title>
					</caption>
					<graphic xlink:href="loquens044_f01" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
			</sec>
			<sec id="S1.3">
				<label>1.3.</label>
				<title>Present study</title>
				<p>Most of the vowels are acoustically differentiated in terms of their first and second formant (<italic>f</italic>
			<sub>1</sub> and <italic>f</italic>
			<sub>2</sub>) values in different languages. In a recent research, <xref ref-type="bibr" rid="b19">Ghaffarvand Mokari and Werner (2016)</xref> found a large overlap in <italic>f</italic>
			<sub>1</sub>–<italic>f</italic>
			<sub>2</sub> space between Azerbaijani /ɯ/ and /œ/ vowels (<xref ref-type="fig" rid="F2">Figures 2</xref> and <xref ref-type="fig" rid="F3">3</xref>). The /ɯ/ and /œ/ vowel are contrastive phonemes in different contexts in Azerbaijani (e.g., /sɯz/ ‘groan’ versus /sœz/ ‘word’).</p>
			<fig id="F2">
					<label>Figure 2.</label>
					<caption>
						<title>Distribution of the Azerbaijani vowels in <italic>f</italic>
			<sub>1</sub>  × <italic>f</italic>
			<sub>2</sub> (Bark) based on production of the 23 female participants in the study by <xref ref-type="bibr" rid="b19">Ghaffarvand Mokari and Werner (2016)</xref>. The ellipses representing two standard deviations from the mean.</title>
					</caption>
					<graphic xlink:href="loquens044_f02" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
				<fig id="F3">
					<label>Figure 3.</label>
					<caption>
						<title>Scatterplot of /ɯ/ and /œ/ vowels based on production of the 23 female participants in the study by <xref ref-type="bibr" rid="b19">Ghaffarvand Mokari and Werner (2016)</xref>. Axes represent <italic>f</italic>
			<sub>1</sub>, <italic>f</italic>
			<sub>2</sub> values in Hz.</title>
					</caption>
					<graphic xlink:href="loquens044_f03" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
		<p>Linear discriminant analysis revealed that <italic>f</italic>
			<sub>1</sub> and <italic>f</italic>
			<sub>2</sub> as predictors fail to accurately classify these two vowels. Further inclusion of <italic>f</italic>
			<sub>0</sub> and duration as the predictors also did not improve the classification percentage. However, the inclusion of <italic>f</italic>
			<sub>3</sub> to the predictors improved the classifications pretty dramatically. It seems these two vowels are more distinct based on <italic>f</italic>
			<sub>3</sub> information (<xref ref-type="fig" rid="F4">Figure 4</xref>). <xref ref-type="bibr" rid="b24">Holt and Lotto (2006)</xref> argue that “if there is not much overlap in a specific acoustic dimension, then that dimension would be very informative about category identity and it would be expected to receive more perceptual weight than the other acoustic dimensions”.</p>
			<fig id="F4">
					<label>Figure 4.</label>
					<caption>
						<title>3D scatterplot of /ɯ/ and /œ/ vowels based on the productions of 23 female participants in the study by <xref ref-type="bibr" rid="b19">Ghaffarvand Mokari and Werner (2016)</xref>. Axes represent <italic>f</italic>
			<sub>1</sub>, <italic>f</italic>
			<sub>2</sub>, and <italic>f</italic>
			<sub>3</sub> values in Hz.</title>
					</caption>
					<graphic xlink:href="loquens044_f04" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
		<p>Based on the hypothesis of <xref ref-type="bibr" rid="b18">Fujimura (1967)</xref> and findings of <xref ref-type="bibr" rid="b7">Chistovich and Lublinskaya (1979)</xref>, since the difference between <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> at least in Azerbaijani /ɯ/ vowel is more than 3.5 Bark (<xref ref-type="bibr" rid="b19">Ghaffarvand Mokari &amp; Werner, 2016</xref>), in this study we assumed <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> as separate perceptual parameters. Studies on how listeners weight perceptual cues in categorization of L1 vowels are limited, especially when vowels are differentiated only by spectral features. We designed the present study in order to explore the perceptual categorization of the two Azerbaijani /ɯ/ and /œ/ vowels. We first aim to find out how listeners weight <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> in discriminating the /ɯ/–/œ/ pair. We aim to find whether <italic>f</italic>
			<sub>3</sub> is an important perceptual cue in the discrimination of this pair or not. We are specifically interested in observing how listeners weigh different acoustic dimensions of these vowels and if they use different cue-weighting strategies in their discrimination.</p>
			</sec>
		</sec>
		<sec id="S2">
			<label>2.</label>
			<title>METHOD</title>
			<sec id="S2.1">
				<label>2.1.</label>
				<title>Participants</title>
				<p>Participants were 10 male and 10 female native Azerbaijani speakers in Tabriz, north-west of Iran. They were born and grown in Tabriz, always used Azerbaijani as the communication language, and reported no history of hearing or other speech problems. They had a mean (SD) age of 30.4 (5.4) years. An informed consent form was obtained from all participants.</p>
			</sec>
			<sec id="S2.2">
				<label>2.2.</label>
				<title>Stimuli</title>
				<p>A 29-year-old male native speaker of Azerbaijani from Tabriz produced several examples of the word /bœl/ ‘divide’ in isolation. All tokens were recorded in a sound-treated room using a ZOOM H6 recorder positioned at approximately 20 cm in front of the speaker. Recordings were at a sampling rate of 44.1 kHz with 16bit resolution. One natural production of the token /bœl/ was selected. The selected token had no sudden changes in formants during the periodic portion of the signal, no changes in fundamental frequency, and no clicks.<italic> </italic>For the resynthesis of the tokens, a periodic portion of the vowel wave form was manually extracted from the /bœl/ token, from the end of the /b/ burst to the last zero crossing of the vowel waveform before the silent gap. The first three formants (<italic>f</italic>
			<sub>1</sub>, <italic>f</italic>
			<sub>2</sub>, and <italic>f</italic>
			<sub>3</sub>) were measured using the standard LPC analysis in Praat (version 6.0.21). The first formant was 475 Hz, the second formant was 1386 Hz, and the third formant was 2273 Hz. The average intensity was 70 dB. We synthesized this token in Praat and created 24 stimuli for the perception experiment. Three sets of tokens were made: (1) by only manipulating the <italic>f</italic>
			<sub>3</sub>, (2) by only manipulating the <italic>f</italic>
			<sub>2</sub>, (3) by manipulating <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub>. For each set, eight spectral steps (equal along a bark scale [1 step = 0.22 Bark for <italic>f</italic>
			<sub>2</sub> and 1 step = 0.19 Bark] for <italic>f</italic>
			<sub>3</sub>) were created (<xref ref-type="fig" rid="F5">Figure 5</xref>).</p>
			<fig id="F5">
					<label>Figure 5.</label>
					<caption>
						<title>Stimuli <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> values.</title>
					</caption>
					<graphic xlink:href="loquens044_f05" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
		<p>We decided to use extreme spectral values within the one standard deviation of the mean for these vowels as absolute exemplars. The values of the two stimuli are based on mean values of the Azerbaijani vowels /œ/ and /ɯ/ in productions of male speakers reported by <xref ref-type="bibr" rid="b19">Ghaffarvand Mokari and Werner (2016)</xref>. The absolute /œ/like instance was the token with the highest <italic>f</italic>
			<sub>2</sub> and the lowest <italic>f</italic>
			<sub>3</sub> (the upper-right corner in <xref ref-type="fig" rid="F5">Figure 5</xref>) and the absolute /ɯ/like instance was the token with the lowest <italic>f</italic>
			<sub>2</sub> and the highest <italic>f</italic>
			<sub>3</sub> in the continuum (the lower-left corner in <xref ref-type="fig" rid="F5">Figure 5</xref>). Additionally, we asked four Azerbaijani native listeners to approve if synthetic tokens were the exemplars of the intended vowels to use in the experiment.</p>
		<p>On each trial of the test in a XAB task, three vowel tokens were played and the listeners were asked to decide whether the first vowel sounded like the second (A) or third (B). The second and third tokens were the most /œ/like and /ɯ/like stimulus, and the first token was one of the 24 stimuli. Participants had to classify each of the 24 stimuli as one of the two absolute exemplars of the vowels. Following <xref ref-type="bibr" rid="b43">Werker and Logan (1985)</xref> and <xref ref-type="bibr" rid="b11">Escudero, Benders, and Lipski (2009)</xref> the interval between the three tokens was set to 1.2 seconds in order to ensure language-specific phonological processing. The order of the presentation of the A and B stimuli was counterbalanced, leading to 48 different XAB trials, which were presented four times each. This way, we ended up with a total of 192 trials.</p>
			</sec>
			<sec id="S2.3">
				<label>2.3.</label>
				<title>Procedure</title>
				<p>The listeners were tested individually in a quiet room by native Azerbaijani speakers who gave all instructions in Azerbaijani. Prior to the experiment, the absolute exemplars of the vowels (the most /œ/like token and the most /ɯ/like token of the stimulus set) were played and participants were asked to pronounce each endpoint and mention three words that include that vowel. This was to ensure that the tokens were easily identifiable as the intended Azerbaijani vowels by native Azerbaijani listeners (<xref ref-type="bibr" rid="b11">Escudero et al., 2009</xref>). The listeners were asked to click on a computer screen displaying the numbers “1”, “2”, and “3”. The number “1” was presented in grey color and was non-clickable. The test was carried out on a PC using Praat. The experiment lasted approximately 20 minutes for each participant. They had a 5minute break in the middle of the experiment. All participants’ accuracy in discrimination of the absolute exem
 plars of either Azerbaijani /œ/ or /ɯ/ was more than 80%, so we did not exclude any of them.</p>
			</sec>
			<sec id="S2.4">
				<label>2.4.</label>
				<title>Analysis</title>
				<p>We performed logistic regression analysis to investigate the listeners’ use of different spectral information. The equation <xref ref-type="disp-formula" rid="form1">(1)</xref> is the example equation used for a model including <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub>.</p>
			<graphic id="form1" xlink:href="loquens044_form1" xmlns:xlink="http://www.w3.org/1999/xlink"/>
			<p>In this equation, <italic>α</italic> is the intercept of the regression model. The coefficients (<italic>β</italic>’s) show how much a one-step difference in one of the predictors makes a change in the log odds of a participant’s response. Hence, according to the suggestions by <xref ref-type="bibr" rid="b31">Morrison (2007,</xref> <xref ref-type="bibr" rid="b32">2009)</xref>, <italic>β</italic> is regarded as participant’s reliance on each of these cues. Following <xref ref-type="bibr" rid="b11">Escudero et al. (2009)</xref> we used equation <xref ref-type="disp-formula" rid="form2">(2)</xref> to compute the relative reliance of the participants on each cue. Values higher than 0.5 mean that <italic>f</italic>
			<sub>2</sub> is weighted heavier than <italic>f</italic>
			<sub>3</sub> and those below 0.5 show that <italic>f</italic>
			<sub>3</sub> is weighted heavier.</p>
			<graphic id="form2" xlink:href="loquens044_form2" xmlns:xlink="http://www.w3.org/1999/xlink"/>
			<p>Also, as mentioned by <xref ref-type="bibr" rid="b11">Escudero et al. (2009)</xref>, polar-coordinate magnitude can be calculated using the logistic regression coefficients, which indicate the boundary crispness. The larger polar-coordinate magnitude indicates a clearer boundary between the two categories (<xref ref-type="bibr" rid="b31">Morrison, 2007</xref>). The polar-coordinate magnitude for the model including only <italic>f</italic>
			<sub>2</sub> was measured as indicated in equation <xref ref-type="disp-formula" rid="form3">(3)</xref>.</p>
			<graphic id="form3" xlink:href="loquens044_form3" xmlns:xlink="http://www.w3.org/1999/xlink"/>
			<p>Based on the results of a logistic regression analysis, it is also possible to find whether an individual cue significantly affects a listener’s responses or not. To this end, we tested whether a logistic regression model that includes a cue as an independent variable predicts the responses significantly better than the null model. For instance, the effect of <italic>f</italic>
			<sub>2</sub> is evaluated through the comparison of the fit of a model with <italic>f</italic>
			<sub>2</sub> as a predictor to the fit of a model with only the intercept. The fit difference between the two models is the Δ<italic>G</italic>
			<sup>2</sup>, which is approximately <italic>χ</italic>
			<sup>2</sup> distributed. The difference in degrees of freedom between the two models are the degrees of freedom of this Δ<italic>G</italic>
			<sup>2</sup>. These results are reported using an <italic>α</italic> level of 0.05 for each participant.</p>
			</sec>
		</sec>
		<sec id="S3">
			<label>3.</label>
			<title>RESULTS</title>
			<p>A series of G<sup>2</sup> comparisons as described in the method section were performed to examine which of the cues were used significantly for vowel categorization. Inclusion of <italic>f</italic>
			<sub>2</sub> significantly improved the fit of the model for 20 out of 20 participants (<italic>p</italic> &lt; 0.05), compared to a model without any independent variable; when only <italic>f</italic>
			<sub>3</sub> was included, the fit of the model significantly improved for 9 participants out of 20 (<italic>p</italic> &lt; 0.05), compared to a model without any factor. <xref ref-type="fig" rid="F6">Figure 6</xref> represents a scatterplot for the coefficients of the regression model with <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> for the 20 participants. <xref ref-type="bibr" rid="b11">Escudero et al. (2009)</xref> mention that “the coefficients of the logistic regression analysis show to what extent a one-step difference in one of the predictors causes a change in the log odds of a participant’s response (p. 457)”.</p>
			<fig id="F6">
					<label>Figure 6.</label>
					<caption>
						<title>Scatterplot of coefficients from the logistic regression analysis that shows the reliance on <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub>.</title>
					</caption>
					<graphic xlink:href="loquens044_f06" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
		<p>As described in 2.4, a cue weighting of 0.5 indicates that the listener weights both cues equally, and a cue weighting higher than 0.5 indicates that the weighting of <italic>f</italic>
			<sub>2</sub> is heavier than that of <italic>f</italic>
			<sub>3</sub>; if it is below 0.5, <italic>f</italic>
			<sub>3</sub> is weighted heavier. <xref ref-type="fig" rid="F7">Figure 7</xref> represents the mean cue weighting for each of the participants. The closer the coefficients of <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> (<xref ref-type="fig" rid="F6">Figure 6</xref>) are to each other, the closer the relative cue weighting is to the center line (both cues weighted equally; <xref ref-type="fig" rid="F7">Figure 7</xref>).</p>
			<fig id="F7">
					<label>Figure 7.</label>
					<caption>
						<title>Relative cue weighting of <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> per participant.</title>
					</caption>
					<graphic xlink:href="loquens044_f07" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
		<p>According to <xref ref-type="fig" rid="F7">Figure 7</xref>, most of the listeners weighted <italic>f</italic>
			<sub>2</sub> heavier, and some listeners weighted <italic>f</italic>
			<sub>3</sub> heavier or both equally. However, there are differences on the amount of weighting of cues among the listeners. Overall, the reliance on <italic>f</italic>
			<sub>2</sub> was much stronger than on <italic>f</italic>
			<sub>3</sub>. Some participants’ relative cue-weighting scores were close to 1 (<italic>f</italic>
			<sub>2</sub> only); however, the three participants who gave more weight to <italic>f</italic>
			<sub>3</sub> did not weight it as strong as the average of participants who gave more weight to <italic>f</italic>
			<sub>2</sub>.</p>
		<p>Finally, we computed the polar-coordinate magnitudes of the two models (only <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>2</sub> + <italic>f</italic>
			<sub>3</sub>) and compared the models’ steepness in categorization boundaries. As mentioned by <xref ref-type="bibr" rid="b31">Morrison (2007)</xref>,</p>
			<disp-quote>
			<p>the contrast coefficient slope in the logistic space is related to the slope of the sigmoidal curve which represents the rate of change from one category to another in the probability space. The size of the contrast coefficient and the corresponding steepness of the steepest tangent to the sigmoidal curve in the probability space are indicators of the crispness of the boundary between the two categories” (p. 229).</p>
			</disp-quote>
			<p><xref ref-type="fig" rid="F8">Figure 8</xref> shows the probability of choosing the /œ/ vowel along the 8 steps. Compared to the model when only <italic>f</italic>
			<sub>2</sub> changes toward the /œ/ vowel, changes of both in <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> make the probability of /œ/ response to jump steeper toward 1.</p>
			<fig id="F8">
					<label>Figure 8.</label>
					<caption>
						<title>Sigmoidal curves in the probability space for contrast coefficient values of model <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>2</sub> + <italic>f</italic>
			<sub>3</sub>.</title>
					</caption>
					<graphic xlink:href="loquens044_f08" xmlns:xlink="http://www.w3.org/1999/xlink"/>
				</fig>
		<p>There was a significance difference between the coefficients of these two models (<italic>t</italic> = -2.03, <italic>p</italic> = 0.05), and inclusion of <italic>f</italic>
			<sub>3</sub> made the curve steeper compared to model with only <italic>f</italic>
			<sub>2</sub> (<xref ref-type="fig" rid="F8">Figure 8</xref>). The mean polar-magnitude values for models were <italic>f</italic>
			<sub>2</sub> + <italic>f</italic>
			<sub>3</sub> = 0.83 &gt; <italic>f</italic>
			<sub>2</sub> = 0.74.</p>
		</sec>
		<sec id="S4">
			<label>4.</label>
			<title>DISCUSSION</title>
			<p>The current study examined perceptual weighting of different acoustic dimensions in perceptual discrimination of Azerbaijani /œ/ and /ɯ/ vowels. Regarding the large overlap in <italic>f</italic>
			<sub>1</sub>–<italic> f</italic>
			<sub>2</sub> vowel space in the production of these two vowels (<xref ref-type="bibr" rid="b19">Ghaffarvand Mokari &amp; Werner, 2016</xref>), this study explored if other cues play a role in their discrimination. To the best of our knowledge, this is the first study of cue weight in the discrimination of native vowels based solely on spectral information. Our results revealed individual differences in the cue weighting for the perception of Azerbaijani /œ/ and /ɯ/ vowels. Although <italic>f</italic>
			<sub>2</sub> was a more important cue for the discrimination of this vowel pair, <italic>f</italic>
			<sub>3</sub> also played a role. Our results of the reliance on different cues revealed that a higher number of listeners mostly relied on <italic>f</italic>
			<sub>2</sub> and some of them relied on <italic>f</italic>
			<sub>3</sub> or on both.</p>
		<p>Overall, one of the important findings in the present study is that <italic>f</italic>
			<sub>2</sub> is still the main cue in the distinction of these two vowels despite their large overlap regarding their <italic>f</italic>
			<sub>2</sub> values. One explanation for this issue would be that listeners use a perceptual vowel-intrinsic normalization process that does not need the information from other vowels. According to <xref ref-type="bibr" rid="b2">Adank et al. (2004)</xref> “vowel-intrinsic normalization models have been considered to be more suitable as models for human vowel perception” because they “can normalize a single vowel from a speaker without information about other vowels from that speaker” (p. 3105).</p>
		<p>These findings are in line with individual differences reported in previous studies. Regarding the discrimination of voiced and voiceless stops, <xref ref-type="bibr" rid="b41">Stevens and Klatt (1974)</xref> report that some listeners relied more on VOT than <italic>f</italic>
			<sub>1</sub> onset frequency, and <xref ref-type="bibr" rid="b21">Haggard et al. (1970)</xref> found that some listeners were more sensitive to <italic>f</italic>
			<sub>0</sub> than VOT in distinguishing voiced and voiceless stops. More recently, <xref ref-type="bibr" rid="b25">Idemaru et al. (2012)</xref> found considerable variability between Japanese listeners’ perceptual weighting of absolute and relative durations in the discrimination of Japanese singleton and geminate stop categories.</p>
		<p>In addition, our results revealed that the categorical boundary was steeper when both <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> were included in the model compared to the model including only <italic>f</italic>
			<sub>2</sub>. This was in line with the study by <xref ref-type="bibr" rid="b22">Hazan and Rosen (1991)</xref>, who observed that listeners’ identification functions were uniformly steep in the full-cue condition.</p>
		<p>Some previous studies have indicated that <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> might be perceptually regarded as one percept (<italic>f</italic>
			<sub>2</sub> prime; <xref ref-type="bibr" rid="b10">Delattre et al., 1952</xref>; <xref ref-type="bibr" rid="b15">Fant &amp; Risberg, 1963</xref>; <xref ref-type="bibr" rid="b16">Fox, Jacewicz, &amp; Chang, 2011</xref>). <xref ref-type="bibr" rid="b7">Chistovich and Lublinskaya (1979)</xref> proposed that close formant peaks are merged into a single perceived spectral prominence. However, other studies suggest that a two-dimensional view on formants cannot explain vowel perception and at least three spectrally prominent regains (corresponding to <italic>f</italic>
			<sub>1</sub>, <italic>f</italic>
			<sub>2</sub>, and <italic>f</italic>
			<sub>3</sub>) are necessary to explain vowel perception (<xref ref-type="bibr" rid="b18">Fujimura, 1967</xref>; <xref ref-type="bibr" rid="b38">Rosner &amp; Pickering, 1994</xref>). There is a need for further studies to investigate whether <italic>f</italic>
			<sub>2</sub> and <italic>f</italic>
			<sub>3</sub> covary in perceptual discrimination of Azerbaijani /œ/ and /ɯ/ vowels and whether they can be merged into one dimension in perception or not.</p>
		<p><xref ref-type="bibr" rid="b17">Francis et al. (2008)</xref> suggest that listeners normally rely on primary cues (e.g., on VOT in the discrimination of English stop voicing contrast) in ideal listening conditions. However, they adjust their cue weighting toward secondary cues under less-than-ideal conditions, for instance when listening to speech in noise or listening to multiple speakers. If <italic>f</italic>
			<sub>2</sub> is the primary cue in the discrimination of the Azerbaijani /œ/ and /ɯ/ vowels, it can be assumed that listeners will use <italic>f</italic>
			<sub>3</sub> cue in noisy and not in ideal conditions.</p>
		<p>Our results also revealed individual differences in categorization gradiency in presence of different cues. In explanation of individual differences in speech perception, <xref ref-type="bibr" rid="b30">Kong and Edwards (2016)</xref> hypothesized that gradiency would be related to general cognitive control. They tested this hypothesis by correlating measures of gradiency with performance on measures of inhibition and task shifting and found little support for this claim. <xref ref-type="bibr" rid="b26">Kapnoula (2016)</xref> also did not find consistent relationships between gradiency and measures of executive function. <xref ref-type="bibr" rid="b25">Idemaru et al. (2012)</xref> speculate that their observed individual cue-weighting pattern can be due to the similar informativeness of the acoustic dimensions and it allows listeners to freely use either source of information, perhaps varying in which information they use across time.</p>
		<p>Future research may look into the relation between production and perception of reliance on spectral cues. One would assume that the individuals who give more weight to <italic>f</italic>
			<sub>3</sub> in discrimination of the Azerbaijani /œ/ and /ɯ/ vowels produce them also with heavier <italic>f</italic>
			<sub>3</sub> differences. In summary, we observed some individual differences in cue-weighting strategies among native listeners. Although there are a few studies on the individual differences in cue weighting, the source of these differences still remains to be discovered in future.</p>
		</sec>
	</body>
	<back>
		<ref-list id="S5">
			<title>REFERENCES</title>
			<ref id="b1">
			<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Abramson</surname>
				<given-names>A. S.</given-names>
			</name>
			<name>
				<surname>Lisker</surname>
				<given-names>L.</given-names>
			</name>
			</person-group>
			<person-group person-group-type="editor">
			<name>
				<surname>Fromkin</surname>
				<given-names>V.</given-names>
			</name>
			</person-group>
			<chapter-title>Relative power of cues: F0 shift versus voice timing</chapter-title>
			<source>Phonetic linguistics: Essays in honor of Peter Ladefoged</source>
			<year>1985</year>
			<fpage>25</fpage>
			<lpage>33</lpage>
			<publisher-name>Academic</publisher-name>
			<publisher-loc>New York</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b2">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Adank</surname>
				<given-names>P.</given-names>
			</name>
			<name>
				<surname>Smits</surname>
				<given-names>R.</given-names>
			</name>
			<name>
				<surname>Hout</surname>
				<given-names>R. van</given-names>
			</name>
			</person-group>
			<article-title>A comparison of vowel normalization procedures for language variation research</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2004</year>
			<volume>116</volume>
			<issue>5</issue>
			<fpage>3099</fpage>
			<lpage>3107</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1795335">https://doi.org/10.1121/1.1795335</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b3">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Ainsworth</surname>
				<given-names>W. A.</given-names>
			</name>
			</person-group>
			<article-title>Duration as a cue in the recognition of synthetic vowels</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>1972</year>
			<volume>51</volume>
			<fpage>648</fpage>
			<lpage>651</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1912889">https://doi.org/10.1121/1.1912889</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b4">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Allen</surname>
				<given-names>J. S.</given-names>
			</name>
			<name>
				<surname>Miller</surname>
				<given-names>J. L.</given-names>
			</name>
			<name>
				<surname>DeSteno</surname>
				<given-names>D.</given-names>
			</name>
			</person-group>
			<article-title>Individual talker differences in voice-onset-time</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2003</year>
			<volume>113</volume>
			<fpage>544</fpage>
			<lpage>552</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1528172">https://doi.org/10.1121/1.1528172</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b5">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Bennett</surname>
				<given-names>D.</given-names>
			</name>
			</person-group>
			<article-title>Spectral form and duration as cues in the recognition of English and German vowels</article-title>
			<source>Language and Speech</source>
			<year>1968</year>
			<volume>11</volume>
			<issue>2</issue>
			<fpage>65</fpage>
			<lpage>85</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/002383096801100201">https://doi.org/10.1177/002383096801100201</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b6">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Bladon</surname>
				<given-names>R. A. W.</given-names>
			</name>
			<name>
				<surname>Lindblom</surname>
				<given-names>B.</given-names>
			</name>
			</person-group>
			<article-title>Modeling the judgment of vowel quality differences</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>1981</year>
			<volume>69</volume>
			<fpage>1414</fpage>
			<lpage>1422</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.385824">https://doi.org/10.1121/1.385824</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b7">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Chistovich</surname>
				<given-names>L. A.</given-names>
			</name>
			<name>
				<surname>Lublinskaya</surname>
				<given-names>V. V.</given-names>
			</name>
			</person-group>
			<article-title>The ‘center of gravity’ effect in vowel spectra and critical distance between the formants: Psychoacoustical study of the perception of vowel-like stimuli</article-title>
			<source>Hearing Research</source>
			<year>1979</year>
			<volume>1</volume>
			<issue>3</issue>
			<fpage>185</fpage>
			<lpage>195</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/0378-5955(79)90012-1">https://doi.org/10.1016/0378-5955(79)90012-1</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b8">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Chládková</surname>
				<given-names>K.</given-names>
			</name>
			<name>
				<surname>Hamann</surname>
				<given-names>S.</given-names>
			</name>
			<name>
				<surname>Williams</surname>
				<given-names>D.</given-names>
			</name>
			<name>
				<surname>Hellmuth</surname>
				<given-names>S.</given-names>
			</name>
			</person-group>
			<article-title>F2 slope as a perceptual cue for the front–back contrast in Standard Southern British English</article-title>
			<source>Language and Speech</source>
			<year>2016</year>
			<volume>60</volume>
			<issue>3</issue>
			<fpage>377</fpage>
			<lpage>398</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/0023830916650991">https://doi.org/10.1177/0023830916650991</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b9">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Crystal</surname>
				<given-names>D.</given-names>
			</name>
			</person-group>
			<source>The Cambridge encyclopedia of language</source>
			<year>2010</year>
			<edition>3</edition>
			<publisher-name>Cambridge University Press</publisher-name>
			<publisher-loc>Cambridge</publisher-loc>
			<publisher-loc>New York</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b10">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Delattre</surname>
				<given-names>P.</given-names>
			</name>
			<name>
				<surname>Liberman</surname>
				<given-names>A. M.</given-names>
			</name>
			<name>
				<surname>Cooper</surname>
				<given-names>F. S.</given-names>
			</name>
			<name>
				<surname>Gerstman</surname>
				<given-names>L. J.</given-names>
			</name>
			</person-group>
			<article-title>An experimental study of the acoustic determinants of vowel color; Observations on one-and two-formant vowels synthesized from spectrographic patterns</article-title>
			<source>Word</source>
			<year>1952</year>
			<volume>8</volume>
			<issue>3</issue>
			<fpage>195</fpage>
			<lpage>210</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/00437956.1952.11659431">https://doi.org/10.1080/00437956.1952.11659431</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b11">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Escudero</surname>
				<given-names>P.</given-names>
			</name>
			<name>
				<surname>Benders</surname>
				<given-names>T.</given-names>
			</name>
			<name>
				<surname>Lipski</surname>
				<given-names>S. C.</given-names>
			</name>
			</person-group>
			<article-title>Native, non-native and L2 perceptual cue weighting for Dutch vowels: The case of Dutch, German, and Spanish listeners</article-title>
			<source>Journal of Phonetics</source>
			<year>2009</year>
			<volume>37</volume>
			<issue>4</issue>
			<fpage>452</fpage>
			<lpage>465</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/j.wocn.2009.07.006">https://doi.org/10.1016/j.wocn.2009.07.006</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b12">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Fant</surname>
				<given-names>G.</given-names>
			</name>
			</person-group>
			<person-group person-group-type="editor">
			<name>
				<surname>Halle</surname>
				<given-names>M.</given-names>
			</name>
			</person-group>
			<chapter-title>On the predictability of formant levels and spectrum envelopes from formant frequencies</chapter-title>
			<source>For Roman Jakobson</source>
			<year>1956</year>
			<fpage>109</fpage>
			<lpage>120</lpage>
			<publisher-name>Mouton</publisher-name>
			<publisher-loc>The Hague</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b13">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Fant</surname>
				<given-names>G.</given-names>
			</name>
			</person-group>
			<source>Acoustic theory of speech production</source>
			<year>1960</year>
			<publisher-name>Mouton</publisher-name>
			<publisher-loc>The Hague</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b14">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Fant</surname>
				<given-names>G.</given-names>
			</name>
			</person-group>
			<source>Speech sounds and features</source>
			<year>1973</year>
			<publisher-name>MIT Press</publisher-name>
			<publisher-loc>Cambridge</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b15">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Fant</surname>
				<given-names>G.</given-names>
			</name>
			<name>
				<surname>Risberg</surname>
				<given-names>A.</given-names>
			</name>
			</person-group>
			<article-title>Auditory matching of vowels with two formant synthetic sounds</article-title>
			<source>STL-Quarterly Progress Status Report</source>
			<year>1963</year>
			<volume>4</volume>
			<issue>4</issue>
			<fpage>7</fpage>
			<lpage>11</lpage>
			</element-citation>
			</ref>
		<ref id="b16">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Fox</surname>
				<given-names>R. A.</given-names>
			</name>
			<name>
				<surname>Jacewicz</surname>
				<given-names>E.</given-names>
			</name>
			<name>
				<surname>Chang</surname>
				<given-names>CY</given-names>
			</name>
			</person-group>
			<article-title>Auditory spectral integration in the perception of static vowels</article-title>
			<source>Journal of Speech, Language, and Hearing Research</source>
			<year>2011</year>
			<volume>54</volume>
			<fpage>1667</fpage>
			<lpage>1681</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1044/1092-4388(2011/09-0279)">https://doi.org/10.1044/1092-4388(2011/09-0279)</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b17">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Francis</surname>
				<given-names>A. L.</given-names>
			</name>
			<name>
				<surname>Kaganovich</surname>
				<given-names>N.</given-names>
			</name>
			<name>
				<surname>Driscoll-Huber</surname>
				<given-names>C.</given-names>
			</name>
			</person-group>
			<article-title>Cue-specific effects of categorization training on the relative weighting of acoustic cues to consonant voicing in English</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2008</year>
			<volume>124</volume>
			<fpage>1234</fpage>
			<lpage>1251</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.2945161">https://doi.org/10.1121/1.2945161</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b18">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Fujimura</surname>
				<given-names>O.</given-names>
			</name>
			</person-group>
			<article-title>On the second spectral peak of front vowels: A perceptual study of the role of the second and third formants</article-title>
			<source>Language and Speech</source>
			<year>1967</year>
			<volume>10</volume>
			<issue>83</issue>
			<fpage>181</fpage>
			<lpage>193</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/002383096701000304">https://doi.org/10.1177/002383096701000304</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b19">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Ghaffarvand Mokari</surname>
				<given-names>P.</given-names>
			</name>
			<name>
				<surname>Werner</surname>
				<given-names>S.</given-names>
			</name>
			</person-group>
			<article-title>An acoustic description of spectral and temporal characteristics of Azerbaijani vowels</article-title>
			<source>Poznań Studies in Contemporary Linguistics</source>
			<year>2016</year>
			<volume>52</volume>
			<issue>3</issue>
			<fpage>503</fpage>
			<lpage>518</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1515/psicl-2016-0019">https://doi.org/10.1515/psicl-2016-0019</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b20">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Ghaffarvand Mokari</surname>
				<given-names>P.</given-names>
			</name>
			<name>
				<surname>Werner</surname>
				<given-names>S.</given-names>
			</name>
			</person-group>
			<article-title>Azerbaijani</article-title>
			<source>Journal of the International Phonetic Association</source>
			<year>2017</year>
			<volume>47</volume>
			<issue>2</issue>
			<fpage>207</fpage>
			<lpage>212</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1017/S0025100317000184">https://doi.org/10.1017/S0025100317000184</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b21">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Haggard</surname>
				<given-names>M.</given-names>
			</name>
			<name>
				<surname>Ambler</surname>
				<given-names>S.</given-names>
			</name>
			<name>
				<surname>Callow</surname>
				<given-names>M.</given-names>
			</name>
			</person-group>
			<article-title>Pitch as a voicing cue</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>1970</year>
			<volume>47</volume>
			<fpage>613</fpage>
			<lpage>617</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1911936">https://doi.org/10.1121/1.1911936</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b22">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Hazan</surname>
				<given-names>V.</given-names>
			</name>
			<name>
				<surname>Rosen</surname>
				<given-names>S.</given-names>
			</name>
			</person-group>
			<article-title>Individual variability in the perception of cues to place contrasts in initial stops</article-title>
			<source>Attention, Perception, &amp; Psychophysics</source>
			<year>1991</year>
			<volume>49</volume>
			<issue>2</issue>
			<fpage>187</fpage>
			<lpage>200</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3758/Bf03205038">https://doi.org/10.3758/Bf03205038</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b23">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Hillenbrand</surname>
				<given-names>J. M.</given-names>
			</name>
			<name>
				<surname>Clark</surname>
				<given-names>M. J.</given-names>
			</name>
			<name>
				<surname>Houde</surname>
				<given-names>R. A.</given-names>
			</name>
			</person-group>
			<article-title>Some effects of duration on vowel recognition</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2000</year>
			<volume>108</volume>
			<fpage>3013</fpage>
			<lpage>3022</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1323463">https://doi.org/10.1121/1.1323463</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b24">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Holt</surname>
				<given-names>L. L.</given-names>
			</name>
			<name>
				<surname>Lotto</surname>
				<given-names>A. J.</given-names>
			</name>
			</person-group>
			<article-title>Cue weighting in auditory categorization: Implications for first and second language acquisition</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2006</year>
			<volume>119</volume>
			<fpage>3059</fpage>
			<lpage>3071</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.2188377">https://doi.org/10.1121/1.2188377</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b25">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Idemaru</surname>
				<given-names>K.</given-names>
			</name>
			<name>
				<surname>Holt</surname>
				<given-names>L. L.</given-names>
			</name>
			<name>
				<surname>Seltman</surname>
				<given-names>H.</given-names>
			</name>
			</person-group>
			<article-title>Individual differences in cue weights are stable across time: The case of Japanese stop lengths</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2012</year>
			<volume>132</volume>
			<fpage>3950</fpage>
			<lpage>3964</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.4765076">https://doi.org/10.1121/1.4765076</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b26">
		<element-citation publication-type="thesis">
			<person-group person-group-type="author">
			<name>
				<surname>Kapnoula</surname>
				<given-names>E. E.</given-names>
			</name>
			</person-group>
			<source>Individual differences in speech perception: sources, functions, and consequences of phoneme categorization gradiency</source>
			<year>2016</year>
			<comment>PhD thesis</comment>
			<publisher-name>University of Iowa</publisher-name>
			</element-citation>
			</ref>
		<ref id="b27">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Kiefte</surname>
				<given-names>M.</given-names>
			</name>
			<name>
				<surname>Kluender</surname>
				<given-names>K. R.</given-names>
			</name>
			</person-group>
			<article-title>Absorption of reliable spectral characteristics in auditory perception</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2008</year>
			<volume>123</volume>
			<fpage>366</fpage>
			<lpage>376</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.2804951">https://doi.org/10.1121/1.2804951</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b28">
		<element-citation publication-type="conf-proc">
			<person-group person-group-type="author">
			<name>
				<surname>Klatt</surname>
				<given-names>D.</given-names>
			</name>
			</person-group>
			<article-title>Prediction of perceived phonetic distance from critical-band spectra: A first step</article-title>
			<conf-name>ICASSP ’82. IEEE International Conference on Acoustics, Speech, and Signal Processing</conf-name>
			<year>1982</year>
			<fpage>1278</fpage>
			<lpage>1281</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1109/ICASSP.1982.1171512">https://doi.org/10.1109/ICASSP.1982.1171512</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b29">
		<element-citation publication-type="conf-proc">
			<person-group person-group-type="author">
			<name>
				<surname>Kong</surname>
				<given-names>E. J.</given-names>
			</name>
			<name>
				<surname>Edwards</surname>
				<given-names>J.</given-names>
			</name>
			</person-group>
			<article-title>Individual differences in speech perception: Evidence from visual analogue scaling and eye-tracking</article-title>
			<conf-name>Proceedings of the International Congress of Phonetic Sciences (ICPhS 17)</conf-name>
			<conf-loc>Hong Kong</conf-loc>
			<conf-date>17–21 August 2011</conf-date>
			<year>2011</year>
			<fpage>1126</fpage>
			<lpage>1129</lpage>
			</element-citation>
			</ref>
		<ref id="b30">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Kong</surname>
				<given-names>E. J.</given-names>
			</name>
			<name>
				<surname>Edwards</surname>
				<given-names>J.</given-names>
			</name>
			</person-group>
			<article-title>Individual differences in categorical perception of speech: Cue weighting and executive function</article-title>
			<source>Journal of Phonetics</source>
			<year>2016</year>
			<volume>59</volume>
			<fpage>40</fpage>
			<lpage>57</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1016/j.wocn.2016.08.006">https://doi.org/10.1016/j.wocn.2016.08.006</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b31">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Morrison</surname>
				<given-names>G.S.</given-names>
			</name>
			</person-group>
			<person-group person-group-type="editor">
			<name>
				<surname>Prieto</surname>
				<given-names>P.</given-names>
			</name>
			<name>
				<surname>Mascaró</surname>
				<given-names>J.</given-names>
			</name>
			<name>
				<surname>Solé</surname>
				<given-names>M.J.</given-names>
			</name>
			</person-group>
			<chapter-title>Logistic regression modelling for first- and second-language perception data</chapter-title>
			<source>Segmental and prosodic issues in Romance phonology</source>
			<year>2007</year>
			<fpage>219</fpage>
			<lpage>236</lpage>
			<publisher-name>John Benjamins</publisher-name>
			<publisher-loc>Amsterdam</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b32">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Morrison</surname>
				<given-names>G. S.</given-names>
			</name>
			</person-group>
			<article-title>L1-Spanish speakers’ acquisition of the English /i/-/ɪ/ contrast II: Perception of vowel inherent spectral change1</article-title>
			<source>Language and Speech</source>
			<year>2009</year>
			<volume>52</volume>
			<issue>4</issue>
			<fpage>437</fpage>
			<lpage>462</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/0023830909336583">https://doi.org/10.1177/0023830909336583</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b33">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Morrison</surname>
				<given-names>G. S.</given-names>
			</name>
			</person-group>
			<person-group person-group-type="editor">
			<name>
				<surname>Morrison</surname>
				<given-names>G.</given-names>
			</name>
			<name>
				<surname>Assmann</surname>
				<given-names>P.</given-names>
			</name>
			</person-group>
			<chapter-title>Theories of vowel inherent spectral change</chapter-title>
			<source>Vowel inherent spectral change</source>
			<year>2013</year>
			<fpage>31</fpage>
			<lpage>47</lpage>
			<publisher-name>Springer</publisher-name>
			<publisher-loc>Berlin</publisher-loc>
			<publisher-loc>Heidelberg</publisher-loc>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1007/978-3-642-14209-3_3">https://doi.org/10.1007/978-3-642-14209-3_3</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b35">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Nearey</surname>
				<given-names>T. M.</given-names>
			</name>
			<name>
				<surname>Kiefte</surname>
				<given-names>M. A.</given-names>
			</name>
			</person-group>
			<article-title>A neural network approach to the dimensionality of the perceptual vowel space</article-title>
			<source>Canadian Acoustics</source>
			<year>2003</year>
			<volume>31</volume>
			<fpage>16</fpage>
			<lpage>17</lpage>
			</element-citation>
			</ref>
		<ref id="b36">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Pols</surname>
				<given-names>L. C. W.</given-names>
			</name>
			<name>
				<surname>Kamp</surname>
				<given-names>L J. Th. van der</given-names>
			</name>
			<name>
				<surname>Plomp</surname>
				<given-names>R.</given-names>
			</name>
			</person-group>
			<article-title>Perceptual and physical space of vowel sounds</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>1969</year>
			<volume>46</volume>
			<fpage>458</fpage>
			<lpage>467</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1911711">https://doi.org/10.1121/1.1911711</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b37">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Raizada</surname>
				<given-names>R. D. S.</given-names>
			</name>
			<name>
				<surname>Tsao</surname>
				<given-names>F.M.</given-names>
			</name>
			<name>
				<surname>Liu</surname>
				<given-names>H.M,</given-names>
			</name>
			<name>
				<surname>Kuhl</surname>
				<given-names>P. K.</given-names>
			</name>
			</person-group>
			<article-title>Quantifying the adequacy of neural representations for a cross-language phonetic discrimination task: Prediction of individual differences</article-title>
			<source>Cerebral Cortex</source>
			<year>2010</year>
			<volume>20</volume>
			<issue>1</issue>
			<fpage>1</fpage>
			<lpage>12</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1093/cercor/bhp076">https://doi.org/10.1093/cercor/bhp076</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b38">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Rosner</surname>
				<given-names>B. S.</given-names>
			</name>
			<name>
				<surname>Pickering</surname>
				<given-names>J. B.</given-names>
			</name>
			</person-group>
			<source>Vowel perception and production</source>
			<year>1994</year>
			<publisher-name>Oxford University Press</publisher-name>
			<publisher-loc>Oxford, UK</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b39">

		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Shultz</surname>
				<given-names>A. A.</given-names>
			</name>
			<name>
				<surname>Francis</surname>
				<given-names>A. L.</given-names>
			</name>
			<name>
				<surname>Llanos</surname>
				<given-names>F.</given-names>
			</name>
			</person-group>
			<article-title>Differential cue weighting in perception and production of consonant voicing</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>2012</year>
			<volume>132</volume>
			<fpage>EL95</fpage>
			<lpage>EL101</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.4736711">https://doi.org/10.1121/1.4736711</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b40">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Stevens</surname>
				<given-names>K. N.</given-names>
			</name>
			</person-group>
			<source>Acoustic phonetics</source>
			<year>1998</year>
			<publisher-name>MIT Press</publisher-name>
			<publisher-loc>Massachusetts</publisher-loc>
			</element-citation>
			</ref>
		<ref id="b41">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Stevens</surname>
				<given-names>K. N.</given-names>
			</name>
			<name>
				<surname>Klatt</surname>
				<given-names>D. H.</given-names>
			</name>
			</person-group>
			<article-title>Role of formant transitions in the voiced-voiceless distinction for stops</article-title>
			<source>The Journal of the Acoustical Society of America</source>
			<year>1974</year>
			<volume>55</volume>
			<fpage>653</fpage>
			<lpage>659</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1121/1.1914578">https://doi.org/10.1121/1.1914578</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b42">
		<element-citation publication-type="book">
			<person-group person-group-type="author">
			<name>
				<surname>Strange</surname>
				<given-names>W.</given-names>
			</name>
			<name>
				<surname>Jenkins</surname>
				<given-names>J. J.</given-names>
			</name>
			</person-group>
			<person-group person-group-type="editor">
			<name>
				<surname>Morrison</surname>
				<given-names>G.</given-names>
			</name>
			<name>
				<surname>Assmann</surname>
				<given-names>P.</given-names>
			</name>
			</person-group>
			<chapter-title>Dynamic specification of coarticulated vowels</chapter-title>
			<source>Vowel inherent spectral change</source>
			<year>2013</year>
			<fpage>87</fpage>
			<lpage>115</lpage>
			<publisher-name>Springer</publisher-name>
			<publisher-loc>Berlin</publisher-loc>
			<publisher-loc>Heidelberg</publisher-loc>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1007/978-3-642-14209-3_5">https://doi.org/10.1007/978-3-642-14209-3_5</ext-link></comment>
			</element-citation>
			</ref>
		<ref id="b43">
		<element-citation publication-type="journal">
			<person-group person-group-type="author">
			<name>
				<surname>Werker</surname>
				<given-names>J. F.</given-names>
			</name>
			<name>
				<surname>Logan</surname>
				<given-names>J. S.</given-names>
			</name>
			</person-group>
			<article-title>Cross-language evidence for three factors in speech perception</article-title>
			<source>Attention, Perception, &amp; Psychophysics</source>
			<year>1985</year>
			<volume>37</volume>
			<issue>1</issue>
			<fpage>35</fpage>
			<lpage>44</lpage>
			<comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.3758/Bf03207136">https://doi.org/10.3758/Bf03207136</ext-link></comment>
			</element-citation>
			</ref>
		</ref-list>
	</back>
</article>
