Você está na página 1de 5

An Acoustic Analysis of Sylheti Phonemes

Amalesh Gope and Shakuntala Mahanta

Phonetics and Phonology Lab, Department of HSS, IIT Guwahati, India


amaleshgope5sept@gmail.com. smahanta@iitg.ernet.in

ABSTRACT [ph] ([por] > [ɸɔr] „r ‟, [phul] > [ɸ l] „flow r‟)


and voiceless velar stop (both aspirated and
This paper examines the acoustic properties of unaspirated) [k] and [kh] ([kali] > [x li] „ink‟, [khal]
Sylheti phonemes. Sylheti is generally regarded as > [x l] „ r in/ h nn l‟ . There was a parallel process
one of the varieties of Bangla. The historical of deaffrication of voiceless alveo-palatal affricate
development of this language witnessed significant [tʃ] and [tʃh] ([tʃa] > [s ] „ ‟, [ ʃhuti] > [suti]
reduction and reconstruction of its phoneme „holi y‟ n oi l o-palatal affricate [dʒ] and
inventory. The phoneme inventory is considerably [dʒ ] ([dʒal] > [z l] „n ‟, [ ʒhal] > [z l] „ pi y‟ and
h

reduced due to the phonological process of together these processes reduced and restructured
deaspiration [+spread glottis], spirantization and the phoneme inventory of the language (Gope &
deaffrication (Gope & Mahanta, 2014). We Mahanta 2014).
conducted an acoustic experiment and measured the Our findings also suggest that even the vowels in
voiced onset time (VOT) of all the voiced stops. The Sylheti have been affected and restructured. There
result of one-way ANOVA did not show any are seven oral vowels in SCB (viz., [i], [e], [æ], [a],
significant interaction among the obstruents in terms [u], [o] and [ɔ]) (Chatterjee 1926, Bhattacharya
of aspiration (p > 0.05, [F (1, 359) = 0.095, p = 1910). However, we have found that the half-open
0.76)]. In a separate experiment, we examined the front vowel [æ] and half-closed back vowel [o] have
acoustic qualities of Sylheti vowels. Results confirm been merged with [e] and [u] respectively and thus
the presence of 5 vowels in Sylheti. A one way reducing the number of vowels to five ([i], [e], [a],
ANOVA confirmed significance effect on vowel [ɔ] and [u] in Sylheti.
quality in terms of duration [F (4, 600) = 57.77, p = This paper demonstrates the phoneme inventory of
0.00] and (first three) formants values. the language under study. Results of two acoustics
experiments have been discussed in this study. The
first experiment is meant to explore the status of the
Keywords: phoneme, voice onset time, duration, underlying property of aspiration [+spread glottis] of
formants. Sylheti obstruents in terms of durational
measurements (VOT). The second production
1. INTRODUCTION experiment is conducted to understand and
determine total number of Sylheti vowels. The
Sylheti is generally regarded as one of the varieties process of spirantization and deaffrication is
of Bangla. There are approximately 10,300,000 discussed with the help of spectrographic evidence
people using Sylheti as their primary language along with adequate examples.
(7,000,000 in Bangladesh) (Lewis et al. 2013). The
most significant properties that distinguish Sylheti 2. SYLHETI PHONEMES: CONSONANTS
from Standard Colloquial Bangla (henceforth SCB)
or other regional varieties is the extensive To understand how the remnant feature [+spread
application of obstruent weakening involving de- glottis], may have affected consonants, we
aspiration, spirantization and deaffrication. conducted an acoustic experiment and measured the
Consequently, the consonant inventory (especially voiced onset time (VOT) of all the voiced stops
the obstruents) of Sylheti exhibit a major reduction present in the language following the theory
and restructuring compared to that of SCB. The loss proposed by Lisker and Abramson (1964).
of (underlying) breathiness contrast [+spread glottis]
of the entire stop series due to the phonological 2.1. Experimental Methodology: VOT
process of obstruent weakening eventually led to
deaspiration of voiced ([ han > an] „paddy‟, [ h > For the acoustic experiment, 18 different words
] „milk‟ n oi l op [ hala > ala] representing each of the target voiced stops
„pl ‟, [m a > ma a] „head‟ . Th process of
h occurring word initially were carefully chosen
spirantization targeted the underlying voiceless (Table 1). As could be seen from the dataset given in
bilabial stop (both aspirated and unaspirated) [p] and the Table 1, the bilabial voiced stop [b] is
represented word-initially in 4 different words (viz., 2.1.2. Acoustic Measurement
[b ] „ r hri i ‟, [b n] „ i ‟, [b l ] „br l ‟ n
[b ri] „hom ‟ . Th other set of words in the Dutta (2007) claimed that voice lead time (VLT)
contrastive pairs (viz., [b ] „ri ‟, [b n] „pr n ‟, (along with the f0 measurements of the following
[b l ] „goo ‟ n [b ri] „h y‟ r pr n the vowel) can successfully lead to a distinction between
(underlying) aspirated voiced stop [bh] which at voiced aspirated and unaspirated counterparts; VLT
some point of history had underlying aspiration, and of voiced aspirate stops will be shorter and hence
appeared as distinct phonemes. The vowel following will reduce the f0 of the following vowel more than
the target stops was an [a] and form (near) minimal the unaspirated voiced stop. Besides, we did not
pairs. Thus, the dental voiced stop [ ] (along with observe any noise like signal following the release in
the underlying aspirated counterpart) is represented the interval that was proposed by Lisker and
in two words, while the retroflex voiced stop [ɖ] and Abramson (1964) as the vital factor to distinguish
the two voiced categories of aspirated and
the velar voiced stop [ɡ] long wi h h ir pir
unaspirated (Figure 1). Hence we decided to
counterparts are represented in 4 different words
consider onset of the release till the offset of the
each. Our major goal was to measure the VOT of
voicing burst for the durational measurements of the
each of the target stops and compare the intrinsic
voiced stop categories.
phonetic variations involving place of articulations
and aspiration.
Figure 1: Waveform of Sylheti words [ɡ ]
„wo n ‟ [ɡh > ɡ]; Th nl rg por ion of h rg
Table 1: Dataset considered for VOT experiment phoneme considered for VOT measurement is
shown along with.

2.1.1. Participants and recording Procedure


2.2. Results and Discussion
Eight native speakers of Sylheti (six male, two
female) were recorded in a quiet environment in A one way ANOVA was performed on the
Dharmanagar district of North Tripura, India. Apart durational values calculated for underlying aspirated
from Sylheti, the speakers were also fluent in Hindi, and unaspirated voiced stops using SPSS software
and English. Speech data was recorded with a Shure (version 20). Altogether 360 tokens were considered
unidirectional head-worn microphone connected to a for acoustic and statistical analysis. Few tokens were
Tascam linear PCM recorder (confirming a constant left out due to either distortion from external noise,
mike-to-mouth distance) via xlr jack. The material or in the presence of noises emanating from fits and
with target stops was displayed on a computer starts in speaking into the microphone. For the
screen. The meaning of each word was written along ANOVA test, Voicing types were kept as categorical
with the sentence frame. Subjects were asked to factor (Independent factor, with two levels viz.;
pronounce each word with natural intonation. To underlying aspiration and unaspiration) and duration
avoid the effect of neighbouring sounds, each word values as variables. As expected, the underlying
was recorded individually with a considerable voiced aspirate stops did not show any difference
amount of gap between two words. Apart from the with their unaspirated counterparts (p > 0.05, [F (1,
target stimuli mentioned in Table 1, an additional 20 359) = 0.095, p = 0.76)]. The duration values of
words were also placed as fillers in the dataset. All individual voiced stop are shown in Figure 2.
the words (along with the target stimuli) were
Figure 2: Average duration of the voiced stop
randomized and presented on three different lists,
consonants with standard deviation as error bars
thus ensuring each stimuli was recorded three times.
Figure 3: Waveform display of the phoneme [ɸ]
as captured in the word [ɸɔr] „r ‟

To observe the interaction among the voiced stop Further, the underlying form of voiceless alveolar
categories a subsequent post-hoc Tukey test was also affricate (both unaspirated and aspirated [tʃ] and
conducted on the data from all the speakers (Table [tʃh]), [tʃir ] > [ ir ] „fl n ri ‟, [k tʃ] > [xas]
3). The findings of post-hoc Tukey test revealed that „gl ‟, [tʃha ] > [s ] „roof‟, [m tʃh] > [m ] „fi h‟,
Sylheti stop categories significantly differ from each and voiced alveolar affricate (both unaspirated and
other in terms of their place of articulation (POA) aspirated [dʒ] and [dʒ h]), [dʒal > [z l] „n ‟, [dʒ hal]
except for the pair [ ] n [ɖ] which could be due to > [z l] „ho ‟, exhibit the process of deaffrication and
less number of tokens considered for the statistical change to alveolar fricatives [s] and [z] respectively.
analysis. As assumed no significant interaction was Table 3 shows the consonant inventory of Sylheti
observed between individual pair of unaspirated
Figure 4: Waveform display of the phoneme [s] as
stops and their aspirated counterparts. No ignifi n
captured in the word [sira] „fl n ri ‟
in r ion w l o ob r b w n [b] n [ ],
how r, h n rlying pir [ ] n [b] iff r
significantly from h o h r wh r h form r w
fo n o b ignifi n ly hor r in r ion h n
h l r. imil rly, n rlying pir [b] n
n pir n l [ ] n n rlying pir [ h] is
observed to be significantly different.

Table 2: Significant matrix for duration of


Sylheti stop consonants Table 3: Consonant inventory of Sylheti

3. SPIRANTIZATION AND DEAFFRICATION 4. ACOUSTIC ANALYSIS OF VOWELS

The underlying forms of voiceless bilabial stop (both To validate the number of vowels and to provide a
aspirated and unaspirated) [p] and [p h] ([por] > [ɸɔr] detailed acoustic analysis of the same, we conducted
„r ‟, [ʃap] > [ʃaɸ] „ n k ‟, [phul] > [ɸ l] „flow r‟, another production experiment and recorded seven
[laph] > [laɸ] „l p/j mp‟) and voiceless velar stop native Sylheti speakers (five male, two female),
(both aspirated and unaspirated) [k] and [kh] ([kali] using the data given in (Table 4). Four different
> [x li] „ink‟, [kak] > [xax] „ row‟, [khali > xali] consonantal contexts such as [bVl], [sVɸ], [xVl] and
„ mp y‟, [mukh] > [mux] „f ‟) have been [zVl] (V being the target vowel) were prepared as
spirantized due to the phonological process of stimuli that represented all the possible vowels in the
consonant weakening. The transformation of language. All the target vowels (embedded in
voiceless bilabial stop [p] to voiceless bilabial different consonantal contexts) were recorded in a
fricative [ɸ] is shown in Figure 3. The enlarged fixed carrier sentence such as [ami cVc xɔi r] „I V
portion of the target phoneme is also shown along i ‟, wh r V i on of h fo r iff r n
with. consonantal environment and V is the target vowel
(see Table 4). Two of the target words (comprising measured in Mel. Since formant frequencies usually
the target vowel) were disyllabic and trisyllabic vary across male and female genders (Peterson and
r p i ly iz., [b l ] „orn m n , n [b lb li] Barney 1952), values for the first two formants (F1
„n m of a bir ‟ , how r, w on i r only h n F2 w r norm liz ing Lob no ‟ 1971
first vowel of those two words for the purpose of normalization procedure and were plotted on an F1 –
analysis. All the words were randomized and F2 plane using NORM (Thomas and Kendell 2007)
repeated thrice. Additional thirty words were also (Figure 5).
used as fillers.
Table 4: Dataset prepared for acoustic analysis Figure 5: Vowel diagram showing average
of the vowels Lobanov normalized formant frequencies of the
first two formants with one standard deviation
ellipses. N= 120 per vowel.

4.1. Results and Discussion: Duration

Duration and formant values of each of the target


vowels were measured using a Praat script (version
5.3.04_win32) after manually labelling each The results of the ANOVA test (conducted on non-
phoneme. A one way ANOVA conducted on the normalized formant values separately for male and
data from all the speakers confirmed significance female) confirmed a strong interaction between
effect of vowel quality on duration [F (4, 600) = vowel types based on the first three formant values
57.77, p = 0.00]. The high vowels [u] (Mean (for both genders) for each vowel. For all the three
duration = 81.04 milliseconds, N = 120) and [i] formants analysed in this study, we noticed a
(Mean Duration = 89.77 milliseconds, N = 120) significant interaction among the vowels (for male
appear to be the shortest. A successive post-hoc speakers: F1: p < 0.05 [(F (4, 300) = 83.89, p =
Tukey test was also conducted to observe the 0.000], F2: p < 0.05 [(F (4, 300) = 17.38, p =
interaction between individual vowel pairs. The 0.000]), F3 (p < 0.05 [(F (4, 300) = 8.56, p = 0.000];
results revealed that the high vowels [i] and [u] do and for female speakers: F1: p < 0.05 [(F (4, 300) =
seem to differ (significantly shorter) from the 233.37, p = 0.000], F2 (p < 0.05 [(F (4, 300) =
remaining three vowels in terms of duration. In all 87.62, p = 0.000]), F3 (p < 0.05 [(F (4, 300) = 20.2,
the other cases, vowel durations were not found to p = 0.000]).
be significantly different from each other. This
pattern was observed for both male and female 5. CONCLUSION
p k r ‟ ; i. ., o of ll h 5 ow l , only h
This paper presents the first comprehensive acoustic
high vowels [i and u] were found to be
study of Sylheti phonemes. With the durational
systematically shorter than the remaining three
measurements, we have shown that the underlying
vowels.
Figure 4: Average duration of Sylheti vowels with
aspirated series do not show any significant
standard deviation as error bars difference both in terms of acoustic signal and VOT
measurements. We did not observe any noise
following the burst of underlying voiced stops (as
shown in Figure 1) which prompted us to consider
the onset of the voicing start till the burst point.
Through acoustic waveforms, we have also
discussed the process of consonant weakening
involving spirantization and deaffrication. An
acoustic analysis of the vowels has also been
conducted and vowel inventories of Sylheti have
been presented. To conclude, the study at large, lays
4.2. Results and Discussion: Formant frequencies
down the foundation of a previously undocumented
The first three formant values of each vowel were language.
calculated at the mid-point and the values were
7. REFERENCES [19] Ohala, Manjari. (1983). Aspects of Hindi Phonology.
Delhi: Motilal Banarsidass.
[1] Abramson, A. S. (1962). The vowels and tones of [20] Qayyum, Abdul Md., and Sultana Razia. (2007).
Standard Thai: Acoustical measurements and Prachin o Madhyajuger Bangla Bhashar Abhidhan
experiments. International Journal of American [Dictionary of Old and Medieval Bengali Language].
Linguistics, 28; 2, part II. Bangla Academy, Dhaka.
[2] Anisuzzaman et.al. (1988). Byabaharik Bangla [21] Rose, P. J. (1987). Considerations on the
Uchcharan Abhidhan [A Dictionary of Standard normalization of the fundamental frequency of
Bengali Pronunciation]. National Institute of Mass linguistic tone. Speech Communication Vol. 10, pp.
Communication, Dhaka. 229-247.
[3] Bhattacharya, Krishna. (1910). Bengali Phonetic [22] Sapir, Edward. (1925). Sound patterns in language.
Reader. CIIL, Mysore. Language (1): 37–51.
[4] Boersma, Paul & Weenink, David (2008). Praat: doing [23] Shahidullah, Md. (1965). Bangladesher Anchalik
phonetics by computer (Version 5.3.04_win32) Bhashar Abhidhan [A Lexicon of Bangladeshi
[Computer program]. Retrieved on December 22, Dialects]. Bangla Academy, Dhaka.
2011, from http://www.praat.org/Bradley 1982, [24] Stevens, kenneth and Helen Hanson. (1995).
[5] Chatterji, Suniti Kumar. (1971). The Origin and Classification of glottal vibration from acoustic
Development of the Bengali Language. London: measurements. In Fujimura and Hirano (eds.), pp.
George Allen and Unwin. 147-170.
[6] Das, Shyamal. (1996). Consonant Weakening in [25] Thomas, E. and Kendall, T. 2007. NORM: The vowel
English and Bangla: A Lexical Phonological normalization and plotting suite. Online version:
Investigation. M. Phil Dissertation, EFLU (Formerly http://ncslaap.lib.ncsu.edu/tools/norm/.
CIEFL), Hyderabad.
[7] Disner, S. (1980). Evaluation of vowels normalization
procedures. Journal of the Acoustical Society of
America Vol. 67, pp. 253-261.
[8] Dutta Indranil. (2009). Acoustics of stop consonants in
Hindi: Voicing, fundamental frequency and spectral
intensity. Verlag Dr. Müller: Saarbrücken.
[9] Dixit, R. Prakash. (1989). Glottal Gestures in Hindi
Plosives. Journal of Phonetics 17, pp. 213-237.
[10] Gope, Amalesh. and Irene, Nisha. V. (2010).
Linguistic Analysis of West Tripura Variety of
Bangla: A Descriptive Study. In C. Sivashanmugam et
al. (Eds.). NewPerspectives in Linguistics (pp. 198-
205). Coimbatore: Department of Linguistic,
Bharathiar University.
[11] Gope, Amalesh and Mahanta, Shakuntala. (2014).
Lexical Tone in Sylheti in TAL-2014 (pp 10-14).
[12] Gordon, Mathew. (1996). The Phonetic Structures of
Hupa. UCLA Working Papersin Phonetics Vol. 93,
pp. 164-187.
[13] Iseli, markus, Yen-Lian Shue and Abeer Alwan.
(2007). Age, sex and vowel dependencies of Acoustic
measures related to the voice source. Journal of the
Acoustical Society of America Vol. 121, pp. 2283-
2295.
[14] Kingston, John. (1985). The phonetics and
phonology of the timing of oral and glottal event. PhD
dissertation, UCLA.
[15] Ladefoged, Peter and Ian Maddieson. (1996). The
sounds of the world’s languages. Malden, MA:
Blackwell.
[16] Ladefoged, Peter. (1996). Elements of Acoustic
Phonetics (2nd Edition). The University of Chicago
Press, Chicago & London.
[17] Ladefoged, Peter. (1971). Preliminaries to linguistic
phonetics. Chicago, IL: University of Chicago Press.
[18] M i off, J.A., “ ino-Tibetan linguistics: present
n f r pro p ”, Ann l R i w of
Anthropology, vol. 20, 469-504, 1991.

Você também pode gostar