Hindustani phonology

Phonology of Hindi and Urdu From Wikipedia, the free encyclopedia

Hindustani is the lingua franca of northern India and Pakistan, and through its two standardized registers, Hindi and Urdu, a co-official language of India and co-official and national language of Pakistan respectively. Phonological differences between the two standards are minimal.

Vowels

The oral vowel phonemes of Hindi according to Ohala (1999:102)
More information Front, Central ...
Hindustani vowel phonemes
Front Central Back
short long short long
Close ɪ ʊ
Close-mid
Open-mid ɛː ə ɔː
Open (æː)
Close

Hindustani natively possesses a symmetrical ten-vowel system.[1] The vowels [ə], [ɪ], [ʊ] are always short in length, while the vowels [aː], [iː], [uː], [eː], [oː], [ɛː], [ɔː] are usually considered long, in addition to an eleventh vowel /æː/ which is found in English loanwords. The distinction between short and long vowels is often described as tenseness, with short vowels being lax, and long vowels being tense.[2] Vowels are somewhat longer before voiced stops than before voiceless stops.[3] Additionally, [ɛ] and [ɔ] occur as conditional allophones of /ə/.

Vowel [ə]

/ə/ is often realized more open than mid [ə], i.e. as near-open [ɐ].[3] It is subject to schwa deletion word-medially in certain contexts.

Vowel [aː]

The open central vowel is transcribed in IPA by either [aː] or [ɑː].

In Urdu, there is further short [a] (spelled ہ, as in کمرہ kamra [kəmra]) in word-final position, which contrasts with [aː] (spelled ا, as in لڑکا laṛkā [ləɽkaː]). This contrast is often not realized by Urdu speakers, and always neutralized in Hindi (where both sounds uniformly correspond to [aː]).[4][5]

Vowels [ɪ], [ʊ], [iː], [uː]

Among the close vowels, what in Sanskrit are thought to have been primarily distinctions of vowel length (that is /i, iː/ and /u, uː/), have become in Hindustani distinctions of quality, or length accompanied by quality (that is, /ɪ, iː/ and /ʊ, uː/).[6] The opposition of length in the close vowels has been neutralized in word-final position, only allowing long close vowels in final position. As a result, Sanskrit loans which originally have a short close vowel are realized with a long close vowel, e.g. śakti (शक्तिشکتی 'energy') and vastu (वस्तुوستو 'item') are [ʃəktiː] and [ʋəstuː], usually not *[ʃəktɪ] and *[ʋəstʊ].[7]

Vowels [ɛ], [ɛː]

The vowel represented graphically as اَے (romanized as ai) has been variously transcribed as [ɛː] or [æː].[8] Among sources for this article, Ohala (1999), pictured to the right, uses [ɛː], while Shapiro (2003:258) and Masica (1991:110) use [æː]. Furthermore, an eleventh vowel /æː/ is found in English loanwords, such as /bæːʈ/ ('bat').[9] Hereafter, اَے (romanized as ai) will be represented as [ɛː] to distinguish it from /æː/, the latter.

In addition, [ɛ] occurs as a conditioned allophone of /ə/ (schwa) within the sequence /əɦə/ (/əɦ/ before the next syllable or word-finally due to schwa deletion).[7] This change is part of the prestige dialect of Delhi, but may not occur for every speaker. Here are some examples of this process:

More information Hindi/Urdu, Transliteration ...
Hindi/UrduTransliterationPhonemicPhonetic
कहना / کہنا "to say"kahnā/kəɦ.nɑː/[kɛɦ.nɑː]
शहर / شہر "city"śahar/ʃə.ɦəɾ/[ʃɛ.ɦɛɾ]
ठहरना / ٹھہرنا "to wait"ṭhaharnā/ʈʰə.ɦəɾ.nɑː/[ʈʰɛ.ɦɛɾ.nɑː]
Close

However, the fronting of schwa does not occur in words with a schwa only on one side of the /ɦ/ such as kahānī /kəɦaːniː/ (कहानीکہانی 'a story') or bāhar /baːɦər/ (बाहरباہر 'outside').

Vowels [ɔ], [ɔː]

The vowel [ɔ] occurs in proximity to /ɦ/ if the /ɦ/ is surrounded on one of the sides by a schwa and on other side by a round vowel (due to Hindustani phonotactics, this generally only occurs in the sequences /əɦʊ/ or /ʊɦə/). It differs from the vowel [ɔː] in that it is a short vowel. For example, in bahut /bəɦʊt/ the /ɦ/ is surrounded on one side by a schwa and a round vowel on the other side. One or both of the schwas will become [ɔ], giving the pronunciation [bɔɦɔt].

Nasalisation of vowels

As in French and Portuguese, there are nasalised vowels in Hindustani. There is disagreement over the issue of the nature of nasalisation (barring English-loaned /æ/ which is seldom nasalised).[9] Masica (1991:117) presents four differing viewpoints:

  1. there are no *[ẽː] and *[õː], possibly because of the effect of nasalization on vowel quality;
  2. there is phonemic nasalization of all vowels;
  3. all vowel nasalization is predictable (i.e. allophonic);
  4. Nasalized long vowel phonemes (/ɑ̃ː ĩː ũː ẽː ɛ̃ː õː ɔ̃ː/) occur word-finally and before voiceless stops; instances of nasalized short vowels ([ə̃ ɪ̃ ʊ̃]) and of nasalized long vowels before voiced stops (the latter, presumably because of a deleted nasal consonant) are allophonic.

Masica supports this last view.[10]

Vowel orthography with diacritics and English approximations

The principal vowel phonemes may be organised as follows to demonstrate the orthographic conventions for vowels.

More information Vowels, IPA ...
Vowels
IPA Hindi ISO 15919 Urdu[11] Approximate English
equivalent
Initial Combining Final Medial Initial
ə [12] a ـہ ـ◌َـ اَabout
ā ـا آ far
ɪ िi ◌ِی ـ◌ِـاِstill
ī ◌ِـیـاِیـfee
ʊ u ◌ُو ـ◌ُـاُbook
ū ◌ُو اُو moon
ē ے ـیـایـmate
ɛː ai ◌َـے ◌َـیـاَیـfairy
ō ◌واوforce
ɔː au ◌َـواَوlot (Received Pronunciation)
ʰ [13] h ھ[13] aspiration of the preceding consonant, as in cake
◌̃ [14] ں ـن٘ـ [15] heavy nasalisation of the preceding vowel, like can't in rapid General American English
[16] [17] homorganic nasal before the succeeding consonant, like jungle or branch, and light vowel nasalisation
Close

Consonants

Hindustani has a core set of 28 consonants inherited from earlier Indo-Aryan ancestors. Supplementing these are two consonants that are internal developments in specific word-medial contexts,[19] and seven consonants originally found in loan words, whose expression is dependent on factors such as status (class, education, etc.) and cultural register (Hindi vs Urdu).

Most native consonants may occur geminate (doubled in length; exceptions are /ɽ, ɽʱ, ɦ/). Geminate consonants are always medial and preceded by one of the interior vowels (that is, /ə/, /ɪ/, or /ʊ/). They all occur monomorphemically except [ʃː], which occurs only in a few Sanskrit loans where a morpheme boundary could be posited in between, e.g. /nɪʃ + ʃiːl/ for niśśīl [nɪˈʃːiːl] ('without shame').[9]

For the English speaker, a notable feature of the Hindustani consonants is that there is a four-way distinction of phonation among plosives, rather than the two-way distinction found in English. The phonations are:

  1. tenuis, as /p/, which is like p in English spin
  2. voiced, as /b/, which is like b in English bin
  3. aspirated, as /pʰ/, which is like p in English pin, and
  4. murmured, as /bʱ/.

The last is commonly called "voiced aspirate", though Shapiro (2003:260) notes that,

Evidence from experimental phonetics, however, has demonstrated that the two types of sounds involve two distinct types of voicing and release mechanisms. The series of so-called voice aspirates should now properly be considered to involve the voicing mechanism of murmur, in which the air flow passes through an aperture between the arytenoid cartilages, as opposed to passing between the ligamental vocal bands.

The murmured consonants are believed to be a reflex of murmured consonants in Proto-Indo-European, a phonation that is absent in all branches of the Indo-European family except Indo-Aryan and Armenian.

Notes
  • ¹Only present in Sanskrit loanwords, predominant in Hindi (and marginally in Urdu as well).
  • ²Only present in Arabic and Persian loanwords, predominant in Urdu.
  • ³ Post-alveolar /ʃ/ is sometimes regarded as a substitution for retroflex /ʂ/ in colloquial pronunciation of Hindi. In Urdu (as well as in Hindi), on the other hand, /ʃ/ is a phoneme coming from Arabic and Persian loanwords.
  • ⁴A non-standard phoneme only found in a few words, or in a few dialects. Otherwise an allophone of /n/ before palatals.
  • Marginal and non-universal phonemes are in parentheses.
  • /ɽ/ is lateral [𝼈] for some speakers.[20] ⟨The /ɽ/ sound is lateral in Dravidian languages, see the voiced retroflex lateral flap
  • /ɽʱ/ is mostly de-aspirated to /ɽ/ in colloquial western Hindi,[21] though retained as a distinct phoneme in eastern Hindi.[22][23][24]
  • /x/, /ɣ/, and /q/ are post-velar.[25]
  • /x/, /ɣ/, /z/, and /q/ are mostly replaced by /kʰ/, /ɡ/, /d͡ʒ/, and /k/ respectively in Hindi, except in the careful speech of educated speakers.[26][27][28] /ʒ/ is found in Urdu and is rarer in Hindi, often being replaced with /z/ (or further by /d͡ʒ/, or even /dʒʱ/ in rarer cases) in the latter; an example of a word containing this sound is aźdahā [əʒ.d̪ə.ɦɑː] (अझ़दहाاژدہا 'dragon').[29][30][31]
  • /ŋ/ mostly only occurs in clusters before velars as in aṅkit, but there are also words like tinkā, ākramaṇkārī, mumkin, tanqīd making it phonemic. Sanskritic loans with ṅ occurring elsewhere are made ṅg as in Sanskrit vāṅmaya being pronounced /ʋaːŋ(ɡ)mɛː/.

Stops in final position are not released, although they continue to maintain the four-way phonation distinction in final position. /ʋ/ varies freely with [v], and can also be pronounced [w]. /r/ is usually flapped or trilled.[32] In intervocalic position, it may have a single contact and be described as a flap [ɾ],[33] but it may also be a clear trill, especially in word-initial and syllable-final positions, and geminate /rː/ is always a trill in Arabic and Persian loanwords, e.g. zarā [zəɾaː] (ज़राذرا 'little') versus well-trilled zarrā [zəraː] (ज़र्राذرہ 'particle').[3] The palatal and velar nasals [ɲ, ŋ] occur only in consonant clusters, where each nasal is followed by a homorganic stop, as an allophone of a nasal vowel followed by a stop, and in Sanskrit loanwords.[19][3] However /n/ + velar clusters also occur, eg. /ʊn.kaː/ making /ŋ/ phonemic. There are native murmured sonorants, [lʱ, rʱ, mʱ, nʱ], eg. nanhā, lhēsnā, kulhāṛī, tumhārā (< ślakṣṇa, ślēṣayati, kuṭhāra, yuṣmē-) there are cases where mbh becomes mh as in sam(b)hālna, also from tatsamas e.g. arhat, but these are considered to be consonant clusters with /ɦ/ in the analysis adopted by Ohala (1999).

The fricative /ɦ/ in Hindustani is typically voiced (as [ɦ]), especially when surrounded by vowels, but there is no phonemic difference between this voiced fricative and its voiceless counterpart [h].

Hindustani also has a phonemic difference between the dental plosives and the so-called retroflex plosives. The dental plosives in Hindustani are laminal denti-alveolar (as in Spanish, Portuguese, etc.), and the tongue-tip must be well in contact with the back of the upper front teeth. The retroflex series is not purely retroflex; it actually has an apico-postalveolar (also described as apico-pre-palatal) articulation, and sometimes in words such as ṭūṭā /ʈuːʈaː/ (टूटाٹوٹا 'broken') it even becomes alveolar (especially in Urdu).[34]

In some Indo-Aryan languages, the plosives [ɖ, ɖʱ] and the flaps [ɽ, ɽʱ] are allophones in complementary distribution, with the former occurring in initial, geminate and postnasal positions and the latter occurring in intervocalic and final positions. However, in Standard Hindi they contrast in similar positions, as in nīṛaj (नीड़जنیڑج 'bird') vs niḍar (निडरنڈر 'fearless').[35]

Allophony of [v] and [w]

Hindustani does not distinguish between [v] and [w], specifically Hindi. These are distinct phonemes in English, but conditional allophones of the phoneme /ʋ/ in Hindustani (written in Hindi or و in Urdu), meaning that contextual rules determine when it is pronounced as [v] and when it is pronounced as [w]. /ʋ/ is pronounced [w] in onglide position, i.e. between an onset consonant and a following vowel, as in dvīp (द्वीप​ دویپ, 'island'), and [v] elsewhere (though it is not universal), as in vrat (व्रत ورت, 'vow'). Native Hindi-Urdu speakers are usually unaware of the allophonic distinctions, though these are apparent to native English speakers (as well as usually to the speakers of the Eastern Indo-Aryan languages, such as Odia, Bengali, Assamese, etc.).[36]

When /ʋ/ is preceded by a consonant which itself is preceded by a vowel (i.e. in the environment VC_), the allophony is non-conditional, i.e. the speaker can choose [v], [w], or an intermediate sound based on personal habit and preference, and still be perfectly intelligible. This is due to ambiguity in the syllabification of such words. For example, advait (अद्वैत ادویت) which is underlyingly /əd̪ʋɛːt̪/, may be syllabified as /ə.d̪ʋɛːt̪/, /əd̪.ʋɛːt̪/, or /əd̪.d̪ʋɛːt̪/. Accordingly, the word can be pronounced equally correctly as [əˈd̪wɛːt̪], [əd̪ˈvɛːt̪], or [əd̪ˈd̪wɛːt̪].[36]

In final clusters, /ʋ/ is [w] when preceded by a more sonorous consonant and [v] otherwise. For example, dvandv (द्वंद्व دوندو, 'pair') is pronounced [d̪wən̪d̪w], but garv (गर्व​ گرو, 'pride') is usually (though not always) pronounced [gəɾv].[36]

External borrowing

Sanskrit borrowing has reintroduced /ɳ/ and /ʂ/ into formal Modern Standard Hindi. They occur primarily in Sanskrit loanwords and proper nouns. In casual speech (as well as in Urdu), they are usually replaced with /n/ and /ʃ/.[9] /ɳ/ does not occur word-initially and has a nasalized flap [ɽ̃] as a common allophone.[19][37]

Loanwords from Persian (including some words which Persian itself borrowed from Arabic or Turkic languages), or even directly from Arabic, introduced six consonants, /f, z, ʒ, q, x, ɣ/. Being Perso-Arabic in origin, these are seen as a defining feature of Urdu, although these sounds officially exist in Hindi and modified Devanagari characters are available to represent them.[38][39] Among these, /f, z/, also found in English, French and Portuguese loanwords, are now considered well-established in Hindi (though not in the rural parts of India, as well as among older generation-speakers); indeed, /f/ appears to be encroaching upon and replacing /pʰ/ even in native (non-Persian, non-English, non-Portuguese) Hindi words as well as many other Indian languages such as Punjabi, Bengali, Gujarati and Marathi, as happened in Greek with phi.[19] This /pʰ/ to /f/ shift also occasionally occurs in Urdu (especially in India).[40] While [z] is a foreign sound, it is also natively found as an allophone of /s/ beside voiced consonants, eg, rasgullā, nasbandī. Similarly, v can get devoiced before voiceless consonants, eg., bevqūf.

The other three Persian loans, /q, x, ɣ/, are still considered to fall under the domain of Urdu, and are also used by some Hindi speakers; however, other Hindi speakers may assimilate these sounds to /k, kʰ, g/ respectively.[27][38][41] The sibilant /ʃ/ is found in loanwords from all sources (Arabic, English, Portuguese, Persian, Sanskrit) and is well-established.[9] Some Hindi speakers (especially those from rural areas, as well as among older generation-speakers) pronounce the /f, z, ʃ/ sounds as /pʰ, dʒ, s/, though these same speakers, having a Sanskritic education, may hyperformally uphold /ɳ/ and /ʂ/.[42][26] In contrast, for native speakers of Urdu, the maintenance of /f, z, ʃ/ is not commensurate with education and sophistication, but is characteristic of all social levels.[41] The sibilant /ʒ/, found in loanwords from Persian, Portuguese, French, and English, is very rare and is considered to fall under the domain of Urdu; although it is officially present in Hindi, many speakers of Hindi assimilate it to /z/, /dʒ/, or even /dʒʱ/.[29][26]

Being the main sources from which Hindustani draws its higher, learned terms, English, Sanskrit, Arabic, and to a lesser extent Persian provide loanwords with a rich array of consonant clusters. The introduction of these clusters into the language contravenes a historical tendency within its native core vocabulary to eliminate clusters through processes such as cluster reduction and epenthesis.[43] Schmidt (2003:293) lists distinctively Sanskrit/Hindi biconsonantal clusters /kj, kr, kl, kʃ, sk, st, sʋ, ʃr, ʃʋ, sn, sp, sm, sj, sʋ, nj, tj, tʋ, dj, dʋ, nj, lj, rj, mj, nʋ, lʋ, rʋ, mʋ, dʒj, dʒʋ, tʃj, tʃʋ, rj, ʂk, ʂʈ, ʂɳ, ʂp, ʂm, ʂj, ʂʋ/, and distinctively Perso-Arabic/Urdu biconsonantal clusters /ʃt, ʃk, ʒd, ʒg, fʃ, ʃf, ft, tf, fr, lx, xl, rf, fr, xf, mt, ms, mz, bl, nz, zf, fz, zq, qz, xr, rx, xt, qt, tq, nf, fn, nx, xn, mx, xm, xʃ, ɣz, rɣ/.

Suprasegmental features

Hindustani has a stress accent, but it is not as important as in English. To predict stress placement, the concept of syllable weight is needed:

  • A light syllable (one mora) ends in a short vowel /ə, ɪ, ʊ/: V
  • A heavy syllable (two moras) ends in a long vowel /aː, iː, uː, eː, ɛː, oː, ɔː/ or in a short vowel and a consonant: VV, VC
  • An extra-heavy syllable (three moras) ends in a long vowel and a consonant, or a short vowel and two consonants: VVC, VCC

Stress is on the heaviest syllable of the word, and in the event of a tie, on the last such syllable. If all syllables are light, the penultimate is stressed. However, the final mora of the word is ignored when making this assignment (Hussein 1997) [or, equivalently, the final syllable is stressed either if it is extra-heavy and there is no other extra-heavy syllable in the word or if it is heavy, and there is no other heavy or extra-heavy syllable in the word]. For example, with the ignored mora in parentheses:[44]

More information Hindi spelling, Urdu spelling ...
Examples of Hindustani stress
Hindi spelling Urdu spelling Romanization Pronunciation Gloss
रेज़गारी ریزگاری rēzgārī [ˈreːz.ɡaː.ri(ː)] small change, coins
समिति سمتی samiti [sə.ˈmɪ.t(ɪ)] committee
क़िस्मत قسمت qismat [ˈqɪs.mə(t)] fate
रोज़ाना روزانہ [roː.ˈzaː.na(ː)] daily
किधर کدھر kidhar [kɪ.ˈdʱə(r)] where, where to
जनाब جناب janāb [dʒə.ˈnaː(b)] sir, mister
असबाब اسباب asbāb [əs.ˈbaː(b)] goods, property
मुसलमान مسلمان musalmān [mʊ.səl.ˈmaː(n)] Muslim
परवरदिगार پروردگار parvardigār [pər.ʋər.dɪ.ˈɡaː(r)] epithet of God
Close

Content words in Hindustani normally begin on a low pitch, followed by a rise in pitch.[45][46] Strictly speaking, Hindustani, like most other Indian languages, is rather a syllable-timed language. The schwa /ə/ has a strong tendency to vanish into nothing (syncopated) if its syllable is unaccented.

See also

References

Bibliography

Related Articles

Wikiwand AI