CONTEXT AFFECTS DANISH L2 CHINESE LEARNERS’ IDENTIFICATION OF POSTALVEOLAR SIBILANTS

Sidsel Rasmussen1,2 and Ocke-Schwen Bohn2

1China Studies and 2Department of English, Aarhus University

[email protected] [email protected]

ABSTRACT . The Mandarin syllabary consists of approx. 400 excluding tones, a significantly This study examines the identification of Standard smaller number than that of other major world Chinese initial by native Danish students languages such as English [5]. In Mandarin, the of Chinese as a foreign language (CFL). Segmental palatal sibilants may be followed by an [i], while perception is known to be affected by neighboring dental and retroflex sibilants are followed by a segments, and in the case of Chinese, certain homorganic syllabic [ɹ̩ ] or [ɻ̩ ] [10], and it consonantal contrasts may be enhanced by the quality has been suggested that this alternation in vowel of the following vowel. We examined how well quality enhances the otherwise perceptually similar intermediate learners of Chinese could apply contrast. [11] found that palatals and dental sibilants (implicit) knowledge of L2 phonology in their were less easily discriminated by native and non- identification of Chinese coronal that are native listeners when both were followed by [i], known to pose challenges. This paper focuses on two whereas sibilant discrimination improved sets of postalveolar sibilants ([tɕ, tɕʰ, ɕ] and [tʂ, tʂʰ, significantly for all listeners when the consonants ʂ]) that are often perceived and produced similarly by preceded [i] and [ɹ̩ ] respectively, in accordance with CFL learners. the phonotactics of . The otherwise Results show a hierarchy of correct identification very similar acoustic characteristics of Mandarin depending on the following vowel: /i/ > /u/ > /a/. We sibilants seem to be more distinct in the /i/ context suggest that learners rely on implicit knowledge of because of the alternation in quality of the phonotactics when perceiving non-native contrasts. nucleus. As for the high, rounded , the alternation only pertains to the palatals, which are Keywords: Nonnative speech perception, Mandarin followed by [y], while dental and retroflex sibilants postalveolars, vowel context effects. are all followed by [u]. All series of sibilants may be followed by a low vowel. We should note that the phonemic status of [tɕ, tɕʰ ɕ] is disputed given the 1. INTRODUCTION predictable alternation of the following high vowels, which means that the palatals never contrast The coronal inventory of Standard Mandarin Chinese minimally with some of the other series. (also just Mandarin) consonants is much larger than Syllables initiated by palatals with a non-high nuclear that of Danish, so L1 Danish learners of Chinese as a vowel are often analyzed as containing a medial high foreign language (CFL) must focus their attention on front segment (e.g. [4]), but this segmentation of the a number of consonantal contrasts that they do not Chinese syllable is sometimes contested as it has been otherwise attend to. Specifically, the three-way suggested that there is not a full segment in the contrast of the Mandarin sibilants ([ts, tsʰ, s], [tɕ, tɕʰ, transition between palatal C and non-high V [6, 7] ɕ] and [tʂ, tʂʰ, ʂ]) is quite rare amongst the world’s The aim of this study was to investigate the role of languages [8], and recent studies have investigated phonotactic knowledge on learners’ perception. how CFL learners acquire this novel contrast [9, 14]. Specifically, we wanted to examine how well (if at In addition to these three sets of , and all) intermediate learners of CFL can make use of aspirated and unaspirated affricates, the coronal predictable patterns in the Mandarin consonant- inventory also includes two alveolar stops [t] and [tʰ]. vowel phonotactic restrictions when they perceive The rich inventory of 11 Mandarin coronals differs coronal obstruents. The identification experiment greatly from the Danish inventory with just six included stimuli of 11 initial Mandarin consonants [t, s coronals [t, t ʰ, s, ɕ, tɕ, tɕʰ]. tʰ, ts, tsʰ, s, tɕ, tɕʰ, ɕ, tʂ, tʂʰ, ʂ] in different vowel As well as learning novel contrasts of L2 phones, contexts, but we limit the scope of this paper to the language learners must also acquire knowledge of L2 results and implications derived for the six

2576 postalveolar sibilants. The two series of postalveolars were used in the identification experiment. The 175 are known to be perceived and produced similarly by stimuli were randomly presented twice, resulting in listeners whose L1 has only a single set of 350 trials in all. The results reported in this paper are postalveolars (e.g. L1 English [3, 12]) and based on identification responses for 18 of the 35 specifically for L1 English CFL learners [12] syllables, i.e. the palatal and retroflex sibilants in suggests that the confusion derives from the fact that three different vowel contexts. Table 1 presents the the place of articulation of English /dʒ, tʃ, ʃ/ lies CV combinations of the selected stimuli. between the two Mandarin series. Previous research suggests similar perceptual problems for L1 Danish Table 1: CV combinations selected as stimuli. listeners [13]. For naïve L1 Danish listeners, palatals C and retroflexes map onto the same L1 category, V tʂ tʂʰ ʂ tɕ tɕʰ ɕ constituting in the terminology of the Perceptual /au/ [au] [au] [au][au] [au] [au] Assimilation Model [1] either a Single-Category /u/ [u] [u] [u] [y] [y] [y] assimilation type or Category-Goodness assimilation /i/ [ ɻ̩̩ ][ ɻ̩̩ ][ ɻ̩̩ ] [i] [i] [i] type, and the ability to successfully discriminate such contrasts can be difficult for learners to acquire. In this study we examine L1 Danish learners’ ability to 2.3. Procedure correctly identify the two sets of postalveolars in Participants first received a short written passage of three different vowel conditions to evaluate if learners Chinese text given in characters, which they can make use of implicit knowledge about L2 transcribed in standard orthography. This was syllables in their perceptual judgments. We expect to ensure familiarity with the pinyin letters as that distinct quality of the high vowels will help response categories. The experiment was conducted learners identify the initial postalveolar consonants in a sound attenuated booth in the speech laboratory more correctly, whereas we expect greater confusion at Aarhus University. Some of the participants were between the two places of articulation in the [au] tested in pairs where two students sat in opposite vowel condition as there are no further cues to the facing directions and ran the experiment from contrasts in the syllable. separate computers over high quality headphones at 2. METHODS the same time. The identification experiment was run in Praat [2], and response options were given on the 2.1. Participants screen in corresponding pinyin letters. The entire experiment included 11 initial coronal obstruents in Fifteen third semester students from the China studies Mandarin, so pinyin letters were all included as response categories. The identification experiment as volunteers. Their ages following presentation and discussion of data will be ranged from 20-25. All participants had studied limited to results from the stimuli with initial [tʂ, tʂʰ, Mandarin at university level for one year, but they ʂ, tɕ, tɕʰ, ɕ] (i.e. ). The five alveolar differed in relevant language and immersion consonants [t, tʰ, ts, tsʰ, s] will also be discussed when experiences. Participants reported no hearing the corresponding orthographic responses were selected as response categories for the postalveolar tokens. There was no time limit on the 2.2. Stimuli participants’ responses, but they were encouraged to respond as fast and as accurately as possible. Stimuli Five female native speakers of Mandarin Chinese were presented with an ITI of 0.5 seconds after the were recorded producing three repetitions each of 35 response was registered. CV syllables in the carrier phrase “我说的是 __” (“What I said was __”). Target syllables were presented to the speakers in standard pinyin 3. RESULTS orthography with the tonal for high level (first) , to allow for potentially uncommon or Tables 2, 3, and 4 show the response matrices for the non-existent syllables. The speakers, aged 23-29, six Mandarin postalveolar sibilants as identified by were familiar with pinyin, and their reading of the Danish intermediate learners. Results are presented stimuli was unproblematic. The first author selected according to the vowel context of the stimuli and the best token from each speaker, which then was discussed separately. Only the low vowel context /au/ verified with 97% accuracy by a native Mandarin has the same vowel quality across the two sibilant listener in a separate identification task. This resulted series, which differ in place of articulation. The in stimuli selection of 5 x 35 = 175 unique tokens that quality of high vowels alternates predictably. In the

2577 tables below, cells of correct responses are Table 4: Identification responses of stimuli in /i/ highlighted, and responses of less than 5% are not condition. included. Alveolar response categories are added in Stimuli rows below the six postalveolars in all three tables so Responses tʂ tʂʰ ʂ tɕ tɕʰ ɕ all incorrect responses >5% can be reported. tʂ 89% 7% 6% tʂʰ 84% 5% 6% Table 2: Identification responses of stimuli in the ʂ 90% /au/ condition (see text). Accurate identification tɕ 81% responses are highlighted. tɕʰ 6% 5% 73% 5% ɕ 77% Stimuli Responses tʂ tʂʰ ʂ tɕ tɕʰ ɕ tʰ 8% tʂ 57% 13% ts 7% tʂʰ 83% 9% tsʰ 6% ʂ 64% 8% tɕ 41% 80% Table 4 above shows the identification matrix for tɕʰ 28% the high, unrounded vowel condition. When followed ɕ 31% 89% by [i] or the syllabic consonant [ɻ̩ ], the two series of tʰ 7% 52% Chinese postalveolar sibilants are generally identified tsʰ 5% 9% with higher accuracy (73% - 90%) than in the two other vowel conditions. Place confusion between the Table 2 shows that correct identification of the six two sets of postalveolars is limited to 6% for some of postalveolars varies considerably in the /au/ the affricate pairs. The fricatives /ɕ/ and /ʂ/ are not condition. The aspirated palatal affricate /tɕʰ/ is confused with each other, but sometimes mistakenly misheard as /tʰ/ in 52% of the trials, and only in 9% heard as their respective homorganic aspirated of the responses misidentified as its retroflex affricate. counterpart /tʂʰ/. Palatal place confusion in the Taken together, the data from the three separate direction of retroflexes is limited (8% - 13%), vowel conditions indicate that Danish learners’ whereas two retroflex tokens are misheard as palatals identification of Mandarin postalveolar sibilants is more frequently (31% - 41%). strongly affected by the following vowel.

Table 3: Identification responses of stimuli in /u/ condition. 4. DISCUSSION

Stimuli Responses tʂ tʂʰ ʂ tɕ tɕʰ ɕ This study revealed a number of expected and tʂ 85% 19% unexpected findings regarding the effect of the tʂʰ 5% 84% 21% following vowel on L1 Danish learners’ identification ʂ 69% 11% of Mandarin postalveolar sibilants. We had expected tɕ 7% 69% 6% a larger percentage of misidentification errors between the two sets of postalveolar sibilants in the tɕʰ 7% 7% 83% 23% /au/ condition because of the shared vowel quality ɕ 53% across the place contrast. This hypothesis was s 6% partially supported with a range of these specific place errors from 8% - 41% in the /au/ condition, 7%- Accurate identification scores of postalveolars 19% in the /u/ condition, and 0 – 6% in the /i/ also vary considerably (53% - 85%) when they condition. precede high rounded vowels as shown in Table 3. Since the vowel quality differs greatly between The greatest source of confusion is found for the retroflex and palatal initiated syllables in the /u/ perception of the fricatives, which are misidentified condition ([u] and [y], respectively), we had not as the homorganic aspirated affricate in 21% and 23% expected misidentification of consonants in this of instances. Confusion between the two postalveolar condition to be as frequent as was the case. Danish series is most notable in the direction of retroflexes, has a phonemic distinction between /u/ and /y/, so L1 with 11-19% of palatal tokens misidentified as their Danish learners of Mandarin are sensitive to the front retroflex counterpart. rounded vowel that occurs in some syllables. Beginning learners of Mandarin may initially be

confused by the representation of this vowel in pinyin

orthography, which is after palatals and <ü> after . All participants in this experiment

2578 had demonstrated perfect pinyin transcription of 5. REFERENCES segments, but we did not examine their pronunciation prior to testing. It is possible that some of the [1] Best, C.T. 1995. A direct realist perspective on cross- participants could have pronounced ([tɕʰy]) as language speech perception. In: Strange, W. (ed.), [tɕʰu], but we find this explanation unlikely Speech Perception and Linguistic Experience: Issues in Cross-language Research. Timonium MD: York Press, considering how frequently this syllable occurs in the 167-200. Mandarin lexicon, e.g. in 去, “to go”. It is more likely [2] Boersma, P., Weenink, D. 2018. Praat: Doing that students were not aware of all the cues that may phonetics by computer (Version 6.0.42). Retrieved assist them in distinguishing between palatals and from www.praat.org retroflexes, as the present experiment was [3] Chang, C. B., Haynes, E. F., Yao, Y., Rhodes, R. 2009. deliberately conducted before the participants had A tale of five fricatives: Consonantal contrast in received any explicit teaching on Mandarin phonetics heritage speakers of Mandarin. University of and phonology. It would be interesting to examine the Pennsylvania Working Papers in Linguistics, 15(1). effect of explicit instruction, and it would be equally [4] Cheng, C. C. 1973. A synchronic phonology of Mandarin Chinese. The Hague: Mouton & Co. interesting to compare the present results with those [5] Duanmu, S. 2006. Chinese (Mandarin): Phonology. In from a group of learners whose L1 does not contrast Brown, K. (ed.) Encyclopedia of Language and /u/ and /y/. One could expect that L1 English learners, Linguistics, Oxford, UK: Elsevier, 351-355. whose L1 does not have an /y/ vowel, would have [6] Duanmu, S. 2007. The phonology of standard Chinese. even greater problems identifying retroflex and (2nd ed). Oxford ; New York: Oxford University Press. palatal consonants correctly in the /u/ condition. [7] Ladefoged, P., & Maddieson, I. 1996. The sounds of In the /i/ condition, confusion between the two the world’s languages. Oxford, OX, UK ; Cambridge, postalveolar series was minimal, which supports the Mass., USA: Blackwell Publishers. claim advanced by e.g. [11] that the contrast is [8] Ladefoged, P., Wu, Z. 1984. Places of articulation: An enhanced by the notably different syllable nuclei [i] investigation of Pekingese fricatives and affricates. J. Phonetics 12(3), 267-278. and [ɻ̩ ]. The overall identification accuracy rates are [9] Lai, Y.-H. 2009. Asymmetry in Mandarin affricate much higher in the /i/ condition, and the highest perception by learners of Mandarin Chinese. Language incidence of confusions is limited to 8% of instances and Cognitive Processes 24, 1265-1285. in which [tɕʰ] is classified as /tʰ/. This error does not [10] Lee-Kim, S.-I. 2014. Revisiting Mandarin “apical seem surprising considering that both [tɕʰi] and [tʰi] vowels”: An articulatory and acoustic study. J. IPA, are licit Mandarin syllables. In the /au/ condition the 44(03), 261–282. misidentification rate of [tɕʰ] for /tʰ/ is even higher [11] Li, M., Zhang, J. 2017. Perceptual distinctiveness (52%), and the existence of the two syllables [tɕʰau] between dental and palatal sibilants in different vowel () and [tʰʲau] () explains why this contexts and its implications for phonological particular rhyme is likely to cause confusion between contrasts. Laboratory Phonology: J. Assoc. Lab. Phonol. 8(1). the alveolar stop and the affricate. [12] Norman, J. (1988). Chinese. Cambridge, New York: We did not include responses for alveolar Cambridge University Press. obstruents in this paper, but it is noteworthy that the [13] Rasmussen, S., Bohn, O.-S. 2017. Perceptual aspirated palatal affricate is not only likely to be assimilation of Mandarin Chinese consonants by native confused with its retroflex counterpart: This study Danish listeners. J. Acoust. Soc. Am. 141(5), 3518. presents support for it being misperceived as the [14] Shao, J. 2015. Acquisition of Mandarin coronal aspirated alveolar stop in one condition. According to sibilants by adult Korean speakers. PhD Dissertation, the literature on Mandarin phonetics (e.g. [15]), the City University of Hong Kong palatals are acoustically most similar to the dental [15] Svantesson, J.-O. 1986. Acoustic analysis of Chinese sibilants [ts, tsʰ, s], so there is reason to expect fricatives and affricates. Lund University Working Papers in Linguistics 25(1), 195–211. confusion of the palatals with several other Mandarin consonants. The Mandarin coronal inventory consists of consonants that L2 learners may find difficult to perceive correctly (e.g. [9, 14]), and some of these ACKNOWLEDGEMENT difficulties are clearly related to the L2 learner’s native consonant inventory. This study further We thank Jonas Villumsen for assisting in the data suggests that learners’ correct identification of collection. Mandarin coronal obstruents, specifically the postalveolar sibilants, indeed depends in large part on the vowel context and on the CV combinations that are phonotactically licit in the syllabary.

2579