Showing posts with label phonology. Show all posts
Showing posts with label phonology. Show all posts

Wednesday, December 11, 2024

More Mabaan pharyngeals

Thomas Anour has posted a number of Bible extracts: Mark 10:13-18, John 1:1-13, and James 4:1-3. Comparing these to a published translation from 2002 (from which he sometimes diverges slightly) and to the anonymous dictionary linked in the previous post makes it possible for a beginner to parse much of the text. No more examples of /ħ/ were heard; but another pharyngeal, /ʕ/, was. This phoneme is absent from the online audio version of this Bible translation, but can be heard clearly in Thomas Anour's pronunciation of at least three frequent words, despite occasional variation, and seems to contrast with the glottal stop /ʔ/, as illustrated by the the last few lines of the following table. While one of the words with /ʕ/ is an Arabic loan, the rest clearly are not.

Unfortunately, I don't know yet where it's coming from. I have yet to find any useful cognates to the words with the pharyngeal in the rest of Nilotic, or even in the meager Jumjum dictionary. "We" corresponds to Nuer <kɔn> and (probably?) Dinka /wɔ̂ɔk/.

English Mabaan
(Anour)
Mabaan
(anon)
Mabaan (Anderson)
and [ʕɔ́sì] ɔci ʔɔ́cé
so that [ʕáŋkàː] aŋ-ka ʔáŋkà
because (< Ar.) [ʕásàan] acaan
where [ʔáŋɛ̀] aŋɛ
quotative particle [ʔàgɪ́] agi ʔàgē
we [ʔɔ̂ːn] ɔɔn ʔɔ̆ɔn

Tuesday, December 10, 2024

Mabaan pharyngeals

The least well documented subgroup of West Nilotic is the Burun group, spoken around the borders between Sudan, South Sudan, and Ethiopia. The largest language in this subgroup is Mabaan, spoken in South Sudan, for which there exists at least one dictionary (available without bibliographic information on Roger Blench's site), and several very interesting articles by Torben Andersen. But we are no longer in the era where a non-field linguist could be content to look at printed sources alone; there is a fair amount of Mabaan content on YouTube, including a channel by a BA-trained linguist and first language speaker of Mabaan, Thomas Anour: Learn Maban, African Language with Thomas Anour. (Like and subscribe, or whatever it is you're supposed to do on YouTube to encourage creators.) Between these, that makes enough material to observe an interesting phonological difference.

In Mabaan as described by Torben Andersen and in the aforementioned anonymous dictionary, /h/ seems to show up only in interjections or loans, and /ħ/ is not mentioned at all. The variety spoken by Thomas Anour, however, features a number of words with initial [ħ] (occasionally varying with [h]). A single cognate in a North Burun language, Mayak, suggest that this is the reflex in his variety of *r, which otherwise becomes a semivowel in Mabaan; more would be desirable.

English Mabaan
(Anour)
Mabaan
(anon)
Mabaan
(Andersen)
Jumjum
(Fadul et al.)
Mayak
(Andersen)
sorghum field (?) <hill> [ħîl] <yielo>
"field for dura grain"
- <yiil>
"field, farm"
-
rat <heeñ> [ħéːɲ] <yyeño> "rat" yiiêɲ-ʌ̀
"~, mouse"
<yiiñ> rii-nit̪
sausage tree <heeṭṭa> [ħétà] <wyeṭṭa>
"pod of ~"
- - -
desert <hong> [ħʌ̂ːŋ] <wɔɔŋ>
"wilderness, desert"
- - -
salmon (sic) <hitta> [ħítàː] - - - -
excuse (Ar. izin) <honda> [ħʌ̀ndá] - - - -

Edit (12/12/2024): The Elenchus comparativus (von Hurter, 1800) records, s.v. "souris" (mouse), <hén> for "Abugonos Burun" vs. <rine> for "J. Kurmuk". This is the only word in the list transcribed with initial h - and the only word on the list corresponding to any of the ones above - but seems sufficient to suggest that this pronunciation is indeed old. Among words with *r, one notes Abugonos <yonga> "meat" and <ímaghi> "blood" (Kurmuk <rin>), which do not support the hypothesis of *r > ħ, but, given the imprecise transcription, do not disprove it either. My thanks to Shuichiro Nakao for sending me a link to this exceptionally early source.

Sunday, March 31, 2019

Final r-cluster metathesis in one child's French

My favourite 4-year-old is doing something very interesting these days with final consonant clusters in his French. Many word-final consonant clusters starting with R get metathesised: parle (speaks) becomes [palʀ] (yet parler "to speak" remains [paʀle]), tourne (turn) becomes [tunʀ], herbe (grass) becomes [ebʀ], ferme (close) becomes [femʀ]. On the other hand, "porte" (door) remains [pɔʀt]; regarde (look!) [ʀəgaʀd]; "force" (strength) [fɔʀs]; "mars" (March) [maʀs], "parc" (park) [paʀk]. Presumably the phenomenon is related to sonority: {l, n, m, b} metathesise, {t, d, s, k} do not. But French allows word-final consonant clusters with falling or rising sonority, and he has no trouble with words like "monstre" (monster) [mõstʀ]. Any idea if this is typical in French first language acquisition?

Nothing of the sort happens in his English or his Arabic. Then again, his English is non-rhotic anyway for some reason, and in Arabic he pronounces /r/ as [ʕ]; French is the only one of his languages where he's got the pronunciation of rhotics more or less sorted.

Friday, February 24, 2017

The Origin of Mid Vowels in Siwi

How does a language with a relatively small vowel system react to pressure from a language with a larger one?

Most northern Berber varieties have a simple four-vowel system: tense /a/, /i/, /u/, vs. lax schwa (/ə/, written e in the official orthography), the latter being mostly predictable and limited to closed syllables. In the eastern and southern Sahara, however, we tend to find slightly larger vowel systems, and it looks very much as though proto-Berber had a rather asymmetrical six-vowel system, close to modern Tuareg but missing /o/: it had tense /a/, /e/, /i/, /u/ vs. lax /ɐ/, /ə/.

Siwi Berber, in western Egypt, has a more symmetrical six-vowel system: tense /a/, /e/, /i/, /o/, /u/ vs. lax /ə/. All of these vowels occur in inherited vocabulary as well as in Arabic loanwords. It is obvious by inspection that, in almost all contexts, *ɐ merged into /ə/. But the distribution of /e/ shows little connection with that of *e: in fact, most instances of proto-Berber *e correspond to Siwi /i/. And the origin of /o/ is not immediately clear at all. How did this happen?

My latest article - written together with Marijn van Putten - proposes some answers. It turns out that proto-Berber */e/ was retained in Siwi only before word-final /n/. Most instances of /e/ and /o/ are found in Arabic loanwords. Within inherited vocabulary, almost all instances of /e/ - and all instances of /o/ - are phonetically conditioned innovations, arising from at least three distinct regular sound changes and one sporadic one. The net effect of this "conspiracy" of sound changes is to extend phonemes otherwise almost entirely restricted to Arabic loans into inherited Berber vocabulary.

If you want the full story, go read our article: The Origin of Mid Vowels in Siwi (published in Studies in African Linguistics 45:1-2 (2016), pp. 189-208).

Sunday, August 09, 2015

Can two kids change Algerian Arabic? (Probably not, but let's see.)

In central Algerian Arabic, feminine nouns are usually marked by a suffix -a, which becomes -ət when possessed. Pronominal possessors are indicated by suffixes, eg -i "my", -u "his". The lax vowel ə cannot occur in open syllables; when the suffix starts with a vowel, this is resolved by dropping it. If doing so would result in a three-consonant cluster, then, in certain cases, the latter is broken up by inserting a new schwa after the first consonant in the cluster, and geminating that consonant: thus jəfn-a جفنة "big bowl" becomes jəffən-t-i جفّنتي "my big bowl". I've been trying to figure out when exactly this happens in the dialect of Dellys, and finding a good deal of variation, especially in the treatment of sonorants: some people (especially but not exclusively the older ones) say səlʕ-t-i سلْعتي "my goods", leaving the cluster intact, while others say səlləʕ-t-i سلّعتي. I was surprised, however, to find two children, 8 and 10-year-old siblings, using a strategy not, as far as I know, used by adults for nouns at all: changing the problematic ə into a. This was confirmed not just by elicitation (zənq-at-i زنقاتي "my alley", ʕənb-at-i عنباتي "my grape", səlʕ-at-i سلعاتي "my goods") but also by sentences produced; thus for bəlɣ-a بلغة "pair of flip-flops":
ənta ʕəndək bəlɣ-at-ək w ana ʕəndi bəlɣ-at-i انتا عندك بلغاتك وانا عندي بلغاتي
"You have your flip-flops and I have my flip-flops."
and for xədm-a خدماتك "work", completely unprompted:
kəmmli xədm-at-ək كمّلي خدماتك
"Finish your work."
which his older brother actually corrected to kəmmli xəddəm-t-ək كمّلي خدّمتك.

Adults' speech furnishes one plausible model for this strategy - not in nouns but in participles. The active feminine participle takes direct object pronoun suffixes, identical to the genitive ones except in the 1st person singular. In such forms, -ət becomes -at before a vowel, rather than dropping the ə: šayf-a شايفة "having seen (f.)", šayf-at-u شايفاتهُ "having seen him (f. subject)". But its extension to nouns is something quite new; neither their parents nor their elder brother nor any adult I've met use such forms.

Most probably, the next time I go to Dellys I'll find these two children using the normal forms and denying they ever spoke this way. Even now, they already use the normal form for body parts which almost always occur possessed: rəqb-a رقبة "neck" becomes rəqqəb-t-i رقّبتي "my neck". But what if this innovation instead spreads among their peers? Most likely it won't: there seems to be little evidence for children initiating language change, notwithstanding the idea's widespread adoption by generative historical linguists, and adults' innovations are much more likely to be maintained or copied (cf. Luraghi 2013, Foulkes and Vihman fc; for a potential counterexample, see Moyna 2009). For that very reason, however, it will be worth keeping an eye on them; potential counterexamples are always interesting.

Sunday, May 10, 2015

How to remember numerals better

In all the debate around "Whorfian" effects of language on cognition, one relatively well-known case has received oddly little attention among linguists, despite being widely discussed by psychologists and popularised by Malcolm Gladwell: the effect of word length on short-term memory (Baddeley et al. 1975). Basically, all other things being equal, it's easier to remember a sequence of short words than a sequence of long words. This suggests that our short-term memory for words (what psychologists confusingly call phonological memory) has a capacity limited by length - specifically, the amount that can be pronounced in about 2 seconds (Schweickert & Boruff 1987). That should suggest, in particular, that numbers presented orally will be easier to remember in a language with short numerals than in one with long numerals. (Note that this affects, among other things, IQ test results, since IQ tests typically include tests of numeral recall.)

Psychologists followed up on this by attempting to test this hypothesis with a number of language pairs (for an overview, see Baddeley (1997). Disclaimer: I'm not a psycholinguist, and the following references are certainly not exhaustive). The best-tested and most consistent result concerns Chinese. Mandarin and Cantonese numerals take shorter to say than English ones, and a number of psychologists have accordingly confirmed that Chinese speakers can remember longer numerals than English speakers (Stigler, Lee, & Stevenson (1986), Hoosain & Salili (1987)), even at 4 years old Chen and Stevenson (1988)), and that this applies even when bilinguals are tested across their two languages (Hoosain 1979). It goes further than that, in fact: Chincotta & Underwood (1997) find that, out of Cantonese, English, Greek, Finnish, Swedish, and Spanish, only Cantonese speakers remember significantly more digits than speakers of other languages - and that this difference disappeared if the subjects were prevented from rehearsing the numbers auditorily by being asked to keep repeating "la-la" while being tested, proving its linguistic nature. The difference ranges around 2 digits, with the exact figure depending on the experiment.

Data for other languages is less clearcut. Welsh numerals take longer to say in isolation than English ones, and Ellis & Hennelly (1986) accordingly found that English-Welsh bilinguals can on average remember longer numerals in English than Welsh. Naveh-Benjamin & Ayres (1986) simultaneously tested the hypothesis for university students in Israel speaking English, Spanish, Arabic, and Hebrew natively (but excluding the digits "seven" and "zero"). They found that the average number of digits recalled was highest in English (7.21), followed by Hebrew (6.51), then Spanish (6.37), and lowest in Arabic (5.77); the ordering by average number of syllables per digit, or by average time taken to read a digit, was English, Spanish, Hebrew, Arabic. However, the difference in number of digits recalled was smaller than predicted by the time taken to read a digit in each language, suggesting that other factors were also relevant.

A proviso is necessary: some recent work, without disputing the differences observed, has made a strong case that they relate not simply to length ( Lovatt, Avons, & Masterson 2000), but crucially to phonological factors (Service 2010, Lethbridge, Hinton & Nimmo 2002). This has been argued for Welsh numerals vs. English ones by Murray & Jones (2002), who find that Welsh digits take longer to say in isolation but actually take less time to say in connnected speech than English ones, and that changes of place of articulation at word boundaries negatively affect memory.

The research is curiously selective in terms of languages examined, and many of the experiments don't control for all possible confounding factors, such as diglossia and social status in the case of Welsh or Arabic. Nevertheless, it does at least seem well-established that speaking Chinese gives a short-term digit memory advantage over speaking major European or Semitic languages. So, if for some reason you regularly need to remember long numerals, and your preferred language doesn't happen to be Chinese, how do you compensate for this handicap?

There are two obvious ways to get around this (assuming you care enough about remembering numerals to want to, which depends very much on your tastes and circumstances). One is to remember the number visually (as a sequence of written digits) or even kinesthetically (as a sequence of typing actions), in which case this particular constraint no longer applies (cf. eg Olsthoorn, Andriga, & Hulstijn 2012). This only helps, however, if you remember numerals better visually or kinesthetically than auditorily, and my impression is that most people don't.

A probably more helpful alternative is to establish a code that lets you turn long numerals into much shorter words by identifying digits with single letters or single phonemes. This solution has a very long history in Arabic and Hebrew, in which each letter of the alphabet can be used to represent a digit: 'a is 1, b is 2, etc. (the first 9 digits are units, the second 10 are tens, and the rest are hundreds). Since short vowels are not letters, the resulting word can be given whatever vowels the user sees fit to give it. A common game of later poets using the Arabic script was to encode the date of their poem within the poem as a chronogram; more practically, Moroccan schoolchildren used to memorise the multiplication tables as a series of meaningless words formed by this encoding (Meakin 1905). Chronograms have been formed using Roman numerals, but for memorisation, at least, they are rather ill-adapted to such a system - think how much padding would be required to turn a number like MDCCCLXXXIII into words.

However, the spread of Hebrew studies in Western Europe following the Renaissance, and the increasing importance of memorising statistics there, encouraged European mnemonists to look for ways of emulating this encoding without having to learn a Semitic language. Doing so at a time when place notation was widely used, they introduced a crucial improvement: each consonant represented a digit in a place notation system, rather than a number in an additive notation system. After various cumulative efforts at improvement, this culminated in the early 19th century with the so-called Major system: 0=s/z, 1=t/d, 2=n, 3=m, 4=r, 5=l, 6=š/ž/č/j, 7=k/g, 8=f/v, 9=p/b, with vowels, semivowels, and laryngeals ignored. To remember 94801 (LACITO's zip code), for example, one would turn it into "professed". This system apparently remains in use among professional mnemonists to this day, despite being virtually unknown to wider society.

Perhaps this is why linguists haven't paid more attention to the word-length effect in the context of the Whorfian debate: it's a clear-cut effect of language on cognition, but not a very profound one, in that it should be fixable by some very simple hacks (or even just by borrowing some one else's numerals). But I'm not aware of any experimental work testing the effect of this particular hack on digit recall...

Sunday, April 06, 2014

Darja notes 3: Diminutive kumquats and affricate phonology

Continuing the Darja theme of my previous posts, I learned a new word today, from a speaker of the traditional dialect of Algiers: تشوينة čwina "kumquat".  This is obviously the diminutive of تشينة čina "orange" (a borrowing from Spanish), just as مشيمشة mšimša "loquat" - another originally Asian fruit - is of مشماش məšmaš "apricot". But its form is a handy clue to the sound system of Algerian Arabic.

Some years ago, Jeffrey Heath wrote a key study of Moroccan Arabic phonology, Ablaut and Ambiguity. Among the questions he tackled was the status of تش č: one phoneme, or two? One way to check is to look at its behaviour in diminutives. Words beginning with two consonants in a row form their diminutives by inserting an i after the two consonants, eg لسان lsan "tongue" > لسيّن lsiyyən "little tongue". Words beginning with one consonant followed by a vowel form the diminutive by replacing the vowel with و w and adding i after it, eg شيخ šix "old man" > شويّخ šwiyyəx "little old man". We thus see from تشوينة čwina that تش č behaves like a single consonant in Algerian Arabic, not like a cluster of two consonants. Since ج j is pronounced as an affricate in the north-central dialect under discussion, this conclusion makes sense. For Morocco, judging by Heath's account, the situation is more ambiguous, and speakers don't really seem sure how to form the diminutive; perhaps the same is true in other parts of Algeria.

Monday, December 09, 2013

wləd/wlid- "boy, son": An irregular development

There's a curious feature I recently noticed about the Arabic of Dellys in Algeria (I can't imagine what took me so long, since it's in my own idiolect as well!). In Morocco and western Algeria "boy" and "son" are both ولْد wəld, corresponding regularly to Classical Arabic وَلَد walad. In Dellys, "boy" is ولد wləd, again corresponding regularly (in Morocco, CaLaC and CaLC, where C is any consonant and L is a sonorant, both end up as CəLC; in central Algeria, the former becomes CLəC, the latter CəLC). But with a possessor – ie, in the sense of "son" – is not wləd, but وليد wlid. You can say وليد خويا wlid xu-ya "my brother's son" or وليدك wlid-ək "your son", but not *wləd xuya or *wəld-ək. It's not obvious how to explain this historically; on the face of it, it looks like a completely irregular development. There are a few other nouns derived from the pattern CaCaC – for instance حنش ħnəš "snake", حبق ħbəq "basil" – but I can't think of any cases offhand which frequently occur in the construct state (that is, with a possessor directly following them). It might be compared to the diminutive, but in present-day Dellys Arabic anyway, the diminutive is وليّد wliyyəd, not wlid.

Has anyone come across a similar phenomenon in any other Arabic variety?

Sunday, August 26, 2012

Sacred Phonology

The artistically fruitful intersection between geometry and mysticism has a fairly well-established label: "sacred geometry". Phonology and mysticism have historically intersected in a similar, but perhaps less familiar, way.

The notion of "place of articulation" predates modern linguistics by millennia, as is obvious from the order of the Indic alphabets. In the Arabic context, while it was first developed by early linguists such as al-Farahidi and Sibawayh, it is probably most commonly studied in the context of Qur'anic recitation, where it provides a cross-check on the pronunciation of consonants. The number of places of articulation used is rather larger than in the Western tradition, allowing a more linear ordering of the consonants, as follows: chest (long vowels ā ī ū), lower throat (glottal ʔ h), mid throat (pharyngeal/epiglottal ʕ ħ), high throat (uvular fricatives x ɣ), back tongue (uvular stop q), velar (k), mid-tongue (palatal/postalveolar j š y), back lateral (ḍ), front lateral (l), apico-alveolar (n), front tongue (r), apico-dental (ṭ d t), sibilants (ṣ s z), interdentals (ð̣ θ ð), labiodentals (f), bilabials (b m w). (I have used odd terminology for some positions in an attempt to reflect divisions not usually made by Western linguists.)

This ordering of consonants by place of articulation, familiar to any religious specialist of the period, gave Ibn Arabi a structure onto which he could map his vision of the cosmic order (see Appendix II of the link). The throat is the seat of the intellective world, ie universal underlying principles; the back of the mouth is the higher realm of imagination; the mouth proper is the celestial spheres, followed in front by the elemental globes; finally, the "progeny", or classes of beings, are at the gap between the teeth and lips. In short, the more contingent something is, the higher up the vocal tract - just as sounds originate at its bottom with air expelled from the lungs, are shaped as they pass through the vocal tract, to finally emerge from the mouth.

The metaphor is reasonably effective as it stands; but its one-dimensionality is somewhat unsatisfactory. Most consonants reflect combinations of articulatory gestures, rather than being elementary movements. For instance, the difference between d and n, for instance, lies not primarily (if at all) in the place of articulation, but rather in whether or not air is allowed to flow out of the nose, and the difference between t and d lies in how the vocal chords are held. Wouldn't it be nicer to have a symbolism for consonants that allowed for compositionality? What would Ibn Arabi have done with Element Theory, for instance?

Sunday, July 29, 2012

Arabic /ē/ gets colloquial: the case of al-Kisā'ī

My description of Khalaf's reading in the previous post applies to all readings transmitted from Ḥamzah ibn Ḥabīb al-Taymī of Kūfa: Khalaf, Khallād, Idrīs al-Ḥaddād, and Isħāq al-Warrāq. There is one other set of readings with final -ē, however: those transmitted from ʕAlī ibn Ḥamza al-Kisā'ī of Kūfa, through his students Abū al-Ḥārith and al-Dūrī. Here's an example (sūrat al-Shams, al-Dūrī's reading):

This set shows two interesting differences for the words examined before:

  • Verbs with medial /ē/ in the Ḥamzah tradition simply have [ā]; contrast Ḥamzah's xēba خاب "he lost" with al-Kisā'ī's xāba. In other words, medially *aya and *awa both become ā, just like in the standard Classical pronunciation.
  • Verbs with final /ā/ in the Ḥamzah tradition have /ē/, just like the ones with /ē/; contrast Ḥamzah's talāhā تلاها "it followed it" with al-Kisā'ī's talēhā. In other words, final *aya and *awa both become ē, whereas original *ā remains ā.

The latter development is phonetically quite counterintuitive - why would *awa become ē, when it didn't even contain any front vowel? But it makes more sense when you look at it on a morphogical rather than phonological level. Arabic has a huge number of final-y verbs, and a much smaller number of final-w verbs. In the rather common 3rd-person perfect forms, they are indistinguishable. This makes it tempting to simplify the system by reducing the differences between the two classes, and in fact practically all modern Arabic dialects have taken this to its logical conclusion and simply conjugate all final-w verbs as if they were final-y: thus Algerian Arabic, for instance, has dʕa دعا, dʕit دعيت, yədʕi يدعي instead of daʕā دعا, daʕawtu دعوت, yadʕū يدعو. What al-Kisā'ī is doing looks like an early step along that road.

You may notice that another characteristic of this reading is also distinctly reminiscent of certain modern colloquials, in particular those of Syria: prepausal feminine -ah ة is pronounced -ih.

Tuesday, July 24, 2012

/ē/, Arabic's fourth long vowel

(Warning: This post assumes some knowledge of Arabic, although you should be able to follow the argument even without it.)

Everyone knows that Classical Arabic has three short vowels (a, i, u) and three long (ā, ī, ū). But is this true of all varieties of Classical Arabic? Listen to this recitation of sūrat al-'Aʕlā:

There are several distinct Qur'ān recitation traditions, thought to reflect (in part) early dialectal variation in pronunciation. The best known are Ḥafṣ (Asia and Egypt) and Warsh (mainly North and West Africa); the recitation above is in one of the more obscure ones, Khalaf (ʕan Ḥamzah). In it, you will notice that words like šē'a “he willed” شاء, tansē “you forget” تنسى, appear with ē where more common pronunciations of Classical Arabic would use ā. But not all cases of ā are pronounced ē: contrast for instance ġuθā'-an “chaff” غثاء, “not” لا. Let's try to figure out what's going on here.

Start with the verbs ending in ā. Verbs which end in ā in the 3rd person masculine singular (“he did”), such as hadā “he guided” هدى, ṣallā “he prayed” صلى, daʕā “he invited” دعا, sajā “it covered with darkness” سجا, divide into two classes in other forms, one ending in y, the other in w: haday-ta “you guided” هديت, ṣallay-ta “you prayed” صليت vs. daʕaw-ta “you invite” دعوت, sajaw-ta “you covered in darkness” سجوت. You have just heard that the former set become hadē, ṣallē. For the latter, we will have to examine different sūras: in the 10th verse of sūrat al-Qamar (1:20) we hear daʕā, and in the 2nd of al-Ḍuħā we hear sajā.

Now ordinary three-letter verbs have the same stem throughout: katab-a “he wrote” كتب vs. katab-ta “you wrote” كتبت. What if the same used to be true of these verbs: *haday-a “he guided” vs. *daʕaw-a “he invited”? (The asterisk means that these are just hypothetical forms.) As it happens, that idea is confirmed if you look at one of Arabic's closest relatives. In Ge'ez, the Semitic classical language of Ethiopia, the cognate verbs are pronounced precisely as reconstructed: with aya (eg ṣallaya “he prayed”) and awa (eg ṣalawa “he roasted”, Arabic ṣalā صلا, ṣalaw-ta صلوت). So if we assume those forms were original, then we can easily see what's going on: in ordinary Classical Arabic both original *aya and *awa end up as ā at the end of a word, but in the Khalaf reading they remain distinct: *aya becomes ē, but *awa becomes ā.

A similar division can be made among verbs with medial ā. Verbs with medial ā in the 3rd person masculine singular, such as zāda “he increased” زاد, ħāqa “he surrounded” حاق, kāna “he was” كان, qāla “he said” قال, divide similarly into two classes in their verbal nouns, one in y, the other in w: zayd زيد, ħayq حيق vs. kawn كون, qawl قول. So we might expect a similar original difference: *zayada, ħayaqa vs. *kawana, *qawala. Sure enough, the pronunciation is as expected. Listen to sūrat al-Baqarah, verses 10 and 11 (about 2:00) and sūrat al-'Anʕām, verse 10 (about 3:00): zēda, ħēqa vs. kāna, qāla. A near-minimal pair is provided by sā'a “he was bad” ساء (sūrat al-Munāfiqūn, v. 2, about 0:50) vs. šē'a “he willed” شاء (already heard in sūrat al-'Aʕlā.)

So – depending on how abstract you are willing to make your representations – this variety of classical Arabic seems to have four long vowel phonemes rather than three. It is also unambiguously more conservative in this respect than the mainstream pronunciation reflected both in the Ḥafṣ reading and in educated standard Arabic, which underscores the philological value of such reading traditions.

(Note: The Qur'ānic Arabic Corpus was useful in preparing this post.)