skip to main |
skip to sidebar
The egregious Sepp Blatter, president of FIFA, tried to defuse the impact of his recent inept remarks on tackling racism by getting the newspapers to print a picture of him in the company of Tokyo Sexwale, the black South African politician.
But how do we pronounce Mr Sexwale’s name? Certainly not ˈseksweɪl.
If you search on-line, you find no authoritative answer and several conflicting pieces of advice from amateurs.
An exchange on reddit went
● Spoiler: Tokyo Sexwale is not pronounced the way it's spelled.
● Yup. As my South African-parented girlfriend immediately pointed out, it's "Seh-tongueclick-wah-leh."
and then● The 'x' in the Sex part is pronounced like a soft 'g' in afrikaans.
Meanwhile the online Telegraph told us firmly Tokyo Sexwale (pronounced seh-wa-le)…
This is one of the names I decided to add to the most recent edition of LPD, so I actually checked it out a few years ago (blog, 3 July 2007).
My initial expectation was that the letter x in his name would stand for the voiceless lateral click, as it does in Xhosa and Zulu, where xoxa ‘tell’ is pronounced ˈǁɔːǁá (or, if you prefer greater explicitness in click symbolism, ˈk͡ǁɔːk͡ǁá).
However, my further researches seemed to suggest that Mr Sexwale’s ethnicity is not Xhosa or Zulu but Venda (or Venḓa — the diacritic indicates a dental, as opposed to alveolar, place of articulation). And in Tshivenḓa the letter x has its IPA value, representing a voiceless velar fricative. So he’d be seˈxwaːle.
The BBC Pronunciation Unit confirmed this. Yes, the IPA for our entry [for Sexwale] indicates a velar fricative. The recommendation is based on the advice of colleagues in Focus on Africa, who, according to our history note from 1993, were adamant that the orthographic 'x' is pronounced as a velar fricative.
(That is indeed also how the orthographic g of Afrikaans is pronounced.)
Conclusion: in English we should call him seˈxwɑːleɪ or, failing that, seˈkwɑːleɪ.
On Friday I said Maybe I’ve just not been keeping my eyes open, but I can’t recall reading any surveys of the prevalence or otherwise of what I would like to call nt-reduction.
One resource I overlooked has now been brought to my attention by Kensuke Nanjo, phonetics editor of the Genius English-Japanese Dictionary, Fourth Edition (2006), in a long email which is worth quoting in extenso. He claims that this “G4” is
the only dictionary that distinguishes [nd] (t-voicing) and [n] (t-deletion) for underlying /nt/ in American English. G4 gives [nd] for carpenter, certainty, into, ninety, seventy, Washington as their second variant in American English while it shows variants without /t/ for other /nt/-words like center, dental, Internet, plenty, twenty, winter, etc. with the label "casual AmE".
Kensuke says that the decisions he made were
based on some books and papers that I'd read and personal communications with American phoneticians, perhaps including the late Becky Dauer, but I'm afraid I don't very well remember where I obtained the data. This distinction ([nd] vs. [n] for /nt/), however, is mentioned in the phonetics/phonology chapter I wrote for the book Ando & Sawada (eds.) English Linguistics: An Introduction (2001), so I obtained the data more than a decade ago.
He further comments
You rightly mention that "it does not happen in the environment of a following stressed vowel, as in intend, contain", but both LPD and G4 record /nt/-reduction for Antarctic, perhaps as a sole (?) exception.
— probably because of the transparent morphology which makes ant#arctic seem like a compound comparable to print#out, in which nt-reducers do reduce nt.
He continues Also, I agree with your comment that "[it doesn't] apply to ntr clusters, as in country. The t can be lost in centre/center but not in central", but G4 gives the variants like "inner"-duce and "inner"-duction for introduce and introduction respectively, with the label "casual AmE". This is based on my own daily observation about American English. Needless to say, this is a case of r-to-schwa metathesis, which triggers /nt/-reduction. In fact, I tried to include as many cases of common metathesis as possible in G4, so it gives the American casual pronunciation "hunnerd" for hundred, a case of both r-to-schwa metathesis and lexically restricted /nd/-reduction (e.g. can'idate, fun'amen'al, kin'a, un'erstand, won'erful).
These nd-reductions of casual speech are very relevant, too. Thanks, Kensuke.
Maybe I’ve just not been keeping my eyes open, but I can’t recall reading any surveys of the prevalence or otherwise of what I would like to call nt-reduction.
By this I mean the loss of t from the cluster nt in intervocalic contexts. This makes winter a possible homophone of winner,
painting a possible rhyme of straining and dental a potential rhyme of kennel. As far as I know this is not found in any kind of British speech, and we think of it as an American or Australian characteristic.
The possible AmE pronunciation of continental as ˌkɑ̃ːʔn̩ˈẽnl̩ is quite strikingly different from the BrE ˌkɒntɪˈnentl̩.
Several qualifications are needed.
• In the kind of AmE I am referring to, winter may possibly have a nasalized tap, thus ˈwɪɾ̃ɚ, rather than the more deliberate plain nasal of winner ˈwɪnɚ. Trager and Smith (1951) refer to this as a ‘flap-release short nasal’, how accurately I am not sure. In any case, a distinction based solely on ɾ̃ vs. n cannot be very robust. I suspect that in reality for many Americans (and Australians) winter and winner can be, and often are, pronounced identically.
• I have the impression that the incidence of nt-reduction is subject to regional variation in the US. It seems more prevalent in the south and west, less so in the north-east. Is this so? Do Canadians ever do it? It is also probably subject to stylistic variation, with unreduced nt more careful and the reduced variant more casual. Has anyone ever investigated its sociolinguistic characteristics?
• The environments in which nt-reduction operates seem to be the same as those for t-voicing. In particular, it does not happen in the environment of a following stressed vowel, as in intend, contain, nor of a following unstressed but strong vowel as in intake; nor does it apply to ntr clusters, as in country. The t can be lost in centre/center but not in central.
• Some words may be special cases, In particular, I have the impression that ninety in AmE is often ˈnaɪndi rather than the expected ˈnaɪnti or ˈnaɪni. Does the same apply to seventy? Are there other exceptional cases?
• Special cases of a different kind are the handful of words in which a similar reduction is found in BrE, namely in London and some other kinds of popular English. For Brits who do this, the t can be lost from twenty and plenty, and from prevocalic went and want (as in went out, wanted), but not from words such as winter or painting.
This posting was triggered by my hearing an Australian golf commentator on TV referring to ðə ˌsevn̩ˈiːnθ the seventeenth (hole). This violates the constraint barring nt-reduction before a stressed vowel, and I suspect would not be possible in AmE.
Furthermore, I wonder whether Australian English has taken nt-reduction direct from AmE, rather than via some British source? And if so, is it the first instance of such a sound change?
Latin h tended to be dropped even in classical times, particularly in the middle of words. Thus nihil ‘nothing’ has an alternative form nīl, and mihi an alternative mī, while dē- plus habeo yields dēbeo ‘I owe’.
In initial position it was more tenacious, though even here by classical times it was only the educated classes who pronounced h. At Pompeii, destroyed 79 CE, there are inscriptional forms such as ic for hic ‘this (m.)’, and conversely hire for ire ‘to go’. In his poem about Arrius, Catullus pokes fun at hypercorrections such as hinsidias for insidias. Even the educated sometimes got confused: the letter h in the regular spelling of humor, humerus, and humidus is apparently unetymological.
The Romance languages inherited no phonetic h from Latin. The h that we pronounce nowadays in English words of Romance or Latin origin reflects a spelling pronunciation: habit, hesitate, horror and for most speakers humo(u)r, humid. As we all know, in various other Latin-derived words we have not restored h despite the spelling: there is no h in heir, hono(u)r, honest. In herb Brits and Americans agree to differ.
I was thinking about this because I have been noticing people pronouncing adhere, adherent, adhesion, adhesive without h, thus əˈdɪə etc. In LPD I give only forms that include h — əd ˈhɪə etc. In this I follow Daniel Jones’s EPD, though I notice that the Cambridge EPD now includes the h-less forms. Rightly so; on reflection, I think they are widespread enough to warrant inclusion, at least for BrE.
I have long been aware of the corresponding h-less pronunciation of abhor, which both LPD and the current EPD (but not the DJ EPD) include.
I don’t think there is any tendency towards a spelling-inspired restoration of h in words with the prefix ex-, as exhaust, exhibit, exhilarate, exhort, which all have -gˈz-. But exhale is a notable exception, always having -ksˈh-, and so sometimes is exhume.
You sometimes encounter the hypercorrect spelling exhorbitant for exorbitant. I can’t say I’ve ever heard the corresponding hypercorrect pronunciation, but presumably it exists.
At the EPSJ conference Takahiro Ioroi presented some statistics about the relative frequency of lexical stress patterns in English words. The pedagogical point was to investigate how far L2 English learners are “exposed to attested patterns in the inputs available”.
Ioroi did this by combining data on stress placement from the Carnegie Mellon University Pronouncing Dictionary with word frequency data from the British National Corpus and a word list from a collection of EFL textbooks for Japanese schools.
In this way he demonstrated that the most frequent exemplar of an initial-stressed disyllable in the school textbooks was people (at 3606 per million), followed by very, other, many and our (sic), while in the BNC it was other (1336 per million) followed by only, also, people and any.
The methodology was irreproachable. But some of Ioroi’s findings demonstrate the truth of the old adage “garbage in, garbage out”.
Let’s not quibble about our (which NSs usually pronounce as a monosyllable).
What about disyllables with final stress? The most frequent one in the textbook corpus was about, which is fair enough. But the most frequent one in the BNC, and second most frequent in the textbooks, comes out as into.
Into? But into is stressed on the first syllable, ˈɪntu, ˈɪntə. It does not have final stress. CMUPD says it does: IH0 N T UW1, which is how they represent ɪnˈtuː. CMUPD is wrong, wrong, wrong.
(In running speech, which is not under consideration here, into may of course lose all stress.)
A useful generalization about English words is that all polysyllables have a primary or secondary lexical stress on either the first or the second syllable. So revolution, for example, has the main stress on the penultimate but on the initial syllable a secondary stress: ˌrevəˈluːʃən. Having supplied stress patterns for several complete dictionary headword lists, I can say that the only exceptions I am aware of are (for some speakers) the two unusual words peradventure and forasmuch. Although they are written as single words, some speakers pronounce them pərədˈventʃə, fərəzˈmʌtʃ, as if they were prepositional phrases, like for a change fərəˈtʃeɪndʒ. (Alternatively, they can be ˌpɜːrədˈventʃə, ˌfɔːrəzˈmʌtʃ, which fit the rule.)
What do Ioroi’s stats tell us about such polysyllables? The most frequent BNC words with lexical stress on neither of the first two syllables are purportedly insufficient and valuation. Again, I am afraid, CMUPD is to blame for supplying wrong information, having forgotten to show secondary stress on the initial syllable of each. (But CMUPD does get the stress pattern of revolution correct.)
For the textbook corpus the results are even odder, since the most frequent such words come out as various proper names, mostly Japanese: Sugihara, Nakamura, Yamagata, Morimoto, Antonelli, which CMUPD shows as having stress only on the penultimate. The fact is that these, too, have initial secondary stress. The incontrovertible evidence for this is the ‘stress shift’ effect when followed by another accented word: ˌSugihara’s ˈwidow (found in this passage).
Again, CMUPD is wrong. GIGO.
A simple question from a Japanese university student: how is that’d pronounced?
My immediate answer was to tell him that it’s ˈðæt əd. I still think that’s the riɡht short answer, but things are actually a little more complicated.
Some relevant variables:
(i) The that element will have a strong vowel, ðæt, only if it is demonstrative, as in that’d be fun, that’d be OK, I don’t think that’d work. If it is a relative pronoun, as in someone that’d been here before, it will almost always be weak, ðət.
(ii) The ’d element may stand not only for would, as in the examples given, but also possibly for had, as in that’d never worked in the past. This makes no difference to the pronunciation.
(iii) There is also the possibility of pronouncing that’d as a monosyllable, ðæd or maybe ðæt, perhaps with further contextual assimilation: that’d be OK ˈðæbbi əʊˈkeɪ.
(iv) for speakers who use t-voicing (esp. AmE), the intervocalic stop/tap in the disyllabic version will be voiced. Here’s a case in point from YouTube.
I can’t find any dictionary entry for that’d that includes pronunciation. Many dictionaries do not even have entries for that’ll and that’s, but leave their pronunciation to be inferred from entries at that and (if you are lucky) ’ll and ’s. LPD does have entries for these contracted forms, though.

The LPD entry at ’d mentions the comparable it’d but not that’d. But now I wonder if I ought to add the possible monosyllabic versions of both.
More generally, in what varieties and styles is it usual to write contracted that’d rather than the usual full that would, that had? I’m really not sure.
There were several interesting papers given at the EPSJ conference just over a week ago in Kochi, and I plan to discuss a few of them over the next few days.
Two of the speakers touched on the use of songs and nursery rhymes in the classroom as pedagogical devices to improve the teaching of pronunciation to EFL students. Both concluded that although they can be valuable they nevertheless need to be handled with caution. This is because the rhythm used in singing is not necessarily identical to the rhythm used by NSs in ordinary speech. (Neither of the two speakers furnished detailed preprints or handouts, so what follows is my own thoughts inspired by their presentations.)
Take the location of stresses. In singing these take the form of the rhythmical beats imposed by the music. Generally speaking, song lyrics reflect lexical stress well: where there’s a lexical stress you get a beat, where there isn’t you don’t. But the correspondence is by no means 100%.
ˈJack and ˈJill went ˈup the ˈhill to ˈfetch a ˈpail of ˈwaˈter
ˈJack fell ˈdown and ˈbroke his ˈcrown and ˈJill came ˈtumbling ˈafˈter.

But in ordinary speech we don’t double-stress water and after. On the other hand we might well stress went and fell.
In particular, rhythmic beats in singing are not a good guide to the deaccentuation of function words. Take this example.
ˈI’m ˈdreaming of a ˈwhite ˈChristmas
ˈJust like the ˈones I used to ˈknow
In these lyrics, since there’s no call for contrastive focus on I’m, in ordinary speech we wouldn’t accent it. (Compare ˈI’m ˈdreaming,| but ˈyou’re aˈwake.) So these lyrics would offer a bad model to those EFL learners who tend to accent pronouns inappropriately.
One speaker got into a terrible muddle with the Burns song Comin’ thro’ the Rye.
Gin a body meet a body
Comin thro’ the rye,
Gin a body kiss a body,
Need a body cry?

For the second body, the music imposes a longer, higher-pitched note on the second syllable than on the first. This led the speaker, if I understood him correctly, to conclude that in the song the word is wrongly stressed, as bɒˈdiː. On the contrary, I would say that it is correctly stressed, and neatly demonstrates the point that in English accent may on occasion be manifested by LOWER pitch than that of a following unstressed syllable, and that in disyllables with a short stressed vowel in the first syllable the second syllable may well be of greater duration than the first.
In any case, the stylized strathspey rhythm of the song is pretty different from the rhythm of ordinary speech. I agree that this song is unsuitable for pedagogical use (except possibly for advanced students), not least because it’s in Scots.
I hope I do not need to add that gin here is pronounced ɡɪn and means ‘if’. Perhaps I ought to add it to LPD.