French – MORPH

Making cuts in the wrong places

September 27, 2023 John Hutchinson Comments 1 comment

When you want to look up a word, how do you go about it? The dictionary is organised by the first letter of the word, so that is what you consider first. And when you want to compare languages, what is the first thing to catch your eye? Again, the first sound. Thus, when looking at a set of words like English fish, father, full, Latin piscis, pater, plenus and Scottish Gaelic iasg, athair, làn, the fact that f- in English corresponds to p- in Latin and zero in Scottish Gaelic spring immediately to our attention, reading as we do from left to right.

Thus, we might presume that the beginning of a word is somehow especially stable, and that sounds which appear at the beginning of a word are a good first indicator of etymology. However, in fact the beginning of a word is not so immutable as you might suppose. Famously, Celtic languages have initial consonant mutations, which alters the initial consonant of a word in regular ways depending on grammatical context. So in Welsh, while ‘Wales’ is Cymru, ‘Welcome to Wales’ is Croeso i Gymru, ‘in Wales’ is yng Nghymru and ‘England and Wales’ is Lloegr a Chymru. This is interesting enough, but not the only way that the start of a word may be altered in languages. Indeed, we don’t even have to leave English to find examples of a different phenomenon that can take place in the history of an individual word.

Let us take a word like adder (the snake specifically, not someone that does addition!). We can look for cognates in closely-related languages, but we are immediately presented with a problem: German Natter, Frisian njirre and Icelandic naðra all seem like they should be related (all being words for ‘snake’), but what’s with this n- at the beginning of the word? Things only get more confusing when we notice words like Latin natrix ‘watersnake’, Welsh neidr or Scottish Gaelic nathair, all again showing an n-. Finally, when we look at Old English we find that the word there is næddre! What’s going on? We know that in general English n- doesn’t do anything particularly strange and it certainly doesn’t just disappear from the beginnings of words, as evidenced by numerous forms like name, night, nest, new, and nine which have had an n- since Proto-Indo-European!

The answer lies in a phenomenon that linguists call ‘rebracketing’. This is a fairly straightforward notion; linguists already make use of brackets to show the internal structure of phrases, thus any change in the structure of the phrase is notated by a change in the arrangement of the brackets. (It will be noted that some authors, including the Oxford English Dictionary, use the term metanalysis instead, but the meaning is the same.)

In the case of adder, the confusion comes from the indefinite article, which in English is a before words beginning with a consonant and an before words beginning with a vowel. Thus, if a word begins with an n-, this can find itself being rebracketed onto the indefinite article: thus [a [nadder]] becomes [a-n [adder]]. And this isn’t the only word where this has happened in English either: thus [a [napron]] (from French napperon) became [a-n [apron]]. On the flipside, the opposite is also found, where the -n from the indefinite article finds itself attached to the front of a word that originally began with a vowel, e.g. [an [ewt]] → [a [n-ewt]] or [an [ekename]] → [a [n-ickname]].

Some of these forms have since become the predominant forms of their respective words, but such is not always the case. For example, uncle derives from a French word oncle, ultimately from Latin avunculus. However, those who are familiar with their Shakespeare will remember the Fool in King Lear, who refers to the title character as ‘nuncle’. Here the reanalysis, rather than from the indefinite article, seems to have been on the basis of possessive pronouns mine and thine, which are particularly frequently used with kind terms: thus [mine [uncle]] becomes [my [nuncle]]. Yet, unlike with the other examples, this has not stuck around, perhaps because the other possessive pronouns (his, her, our, your, their) which would not have motivated this reanalysis; thus the original uncle stuck around and was able to reassert itself.

Nor is English alone in exhibiting these kinds of change. In the adder~nadder case, the same reanalysis has also taken place in Dutch and Low German, also spelt adder in both cases. Similarly, Arabic nāranj was borrowed into Spanish as Naranja, but this underwent rebracketing when it was borrowed into Italian as arancia, and it was from there that the word spread to the rest of Europe, including English orange.

French provides us with an especially interesting example of layered reanalyses in a single word. In Old French, unicorne was reanalysed as beginning with the indefinite article (which is in a sense not incorrect: the literal meaning of the word is ‘one-horn’ and ‘one’ is the source of the French indefinite article, as well as indefinite articles in general cross-linguistically). This left a form icorne, which would contract with the definite article, giving l’icorne ‘the unicorn’. However, at some point, this contracted form with the article came to be reanalysed as the base of the noun itself, with the result that licorne is now simply the French for ‘unicorn’, leading to constructions such as la licorne ‘the unicorn’ where a historical definite article appears ‘doubled up’!

Some of the most complex cases of rebracketing can be found in Scottish Gaelic. Here we have a number of potential sources of rebracketing, both because the definite article changes depending on the following noun and because of the interaction of the definite article and the mutation system.

Firstly, with vowel-initial masculine noun the definite article prefixes a t- e.g. eun ‘bird’ but an t-eun ‘the bird’. Unsurprisingly, based on the examples we have seen above, this prefixed t- has in many cases become attached to the noun. Interestingly this is particularly common in loanwords from Old Norse, such as talla ‘hall’ from hǫll, tòb ‘small bay’ from hóp (òb is also common) and tolm ‘small islet’ from holmr, as well as other loans such as taigeis ‘haggis’ and tobha ‘hoe’ from English.

In a similar vein, one of the components of consonant mutation is Scottish Gaelic is that an f sound disappears (though is still written as fh). As a result, a larger number of words that began with vowels in Old Irish have acquired an f- in Scottish Gaelic, e.g. áinne ‘ring’, uar ‘cold’ and íaru ‘squirrel’ have become fáinne, fuar and feòrag respectively, as if an áinne uar ‘the cold ring’ was really an fháinne fhuar. Many of the words have undergone the same kinds of changes in Irish and Manx, though not all languages agree on which (e.g. Irish also has fáinne and fuar but iora respectively).

And, as in English, words that begin with n- can find this consonant being rebracketed as part of the article an. However, once this n- has been rebracketed, this now vowel-initial word can undergo the same kinds of mutation-based reshaping as an originally vowel initial word. Perhaps the most extreme example of this is ‘nettle’, which was nenaid in Old Irish, but in Scottish Gaelic can be (depending on who you ask) any of neanntag, eanntag (with the n- rebracketed away), feanntag (with the f- appended by lenition reversal) and deanntag (where the d- is apparently a hypercorrective reversal of a process of nasalisation in the Northwestern dialects)!

A bed of nettles — neanntag, eanntag, feanntag or deanntag?

So, when searching around for a word in a dictionary or an old text, be cautious; simply looking for the first consonant to give you a clue might be misleading when taken out of context. Furthermore, instances like these make clear that language is primarily a spoken phenomenon and the kinds of changes that we see reflect that: while in a written text the different between a newt and an ewt is obvious, in spoken language the question of where one word ends and the nexts begins is not so straightforward as a casual glance at a dictionary might suggest. Perhaps this should then make us ponder further how much written language is a direct reflection of spoken language versus being at least partially arbitrary choices made by the writers.

Royal Rules on Rich Rectors

March 15, 2023 John Hutchinson Comments 0 Comment

Last month, the Guildford Shakespeare Company put on a production of Richard II, a fascinating tale of political strife and the perils of having a leader lacking in competence when the country is in crisis. Sound familiar? In any case, this got me thinking about the name Richard and its many etymological links.

First with the name Richard. It’s borrowed from French, but it didn’t start there. In fact it is one of a number of French words that was borrowed from Germanic, deriving from Frankish *Rīkahard, meaning ‘hard/brave king’. This also gives modern German Richard and through the travels of the Goths and Vandals also made its way into Spanish as Ricardo and Italian as Riccardo. The first part of this name, the *rīk- ‘ruler’ part, in other derivations also gives words like German Reich and Dutch rijk, both meaning ‘empire’ or ‘kingdom’, which in English is also found as the ‘domain, kingdom’ suffix -ry, as in Jewry ‘the Kingdom of the Jews’. As different derivation again gives us English rich, something you’d rather expect a king to be. As a component of names it is ubiquitous in Germanic, such as in Old English Godric ‘God(ly) king’, Wulfric ‘Wolf-king’ and Theodric ‘King of the people’. This last one turns up in German as Dietrich and, again courtesy of the Franks, through French Thierry comes into English as Terry (see also my previous post on the Germans for more on this Theod-).

But it is not only Germanic languages that have this root. Indeed, some form of it crops up across the Indo-European language family, usually meaning something like ‘king’ or ‘ruler’. In Celtic (from which Germanic likely borrowed the rīk- words) we find e.g. rí in Irish and rhi in Welsh, both meaning king. In Gaulish, rulers such as Vercingetorix and Ambiorix had an earlier form –rix it as part of their name, and in a reduced form we find the same in the Welsh surname Tudor, originally meaning ‘ruler of the people’ and thus cognate with Theodric/Dietrich/Terry.

In Latin too we find rēx, again meaning ‘king’ or ‘ruler’. This form survives as such in many modern Romance languages, for example Spanish rey and French roi. We also get two separate adjectives in English: regal from Latin and royal from French. Further afield, we find this word cropping up as far away as India, in the form of Sanskrit rāja, once again a ‘king’ word, as well as rāṣṭrá, a ‘kingdom’.

All of these forms can be traced back to a form in Proto-Indo-European (the reconstructed ancestor of all of these languages), which we represent as *h3rḗǵs. In the terminology of Indo-European studies this is an ‘athematic root noun’, meaning a short root without additional derivational suffixes onto which inflectional endings such as the nominative singular *-s are suffixed directly, rather than having an additional ‘theme vowel’ *-o inbetween. As with many such forms in Proto-Indo-European, when we isolate the root itself, *h3reg-, which probably meant something like ‘stretch out the arm, direct’, we can find even more related derivations.

Adding a thematic vowel *-e/o- we get a verb which shows up in Latin as regō ‘rule, govern, direct’, along with an array of derived nouns which we have inn English. We have the agent noun rector, the instrument noun rule (from a French reflex of Latin rēgula) and the abstract noun regimen. Additionally, we have prefixed verbs such as dīrigō, ērigō and corrigō, which through their respective supine forms dīrēctum, ērēctum and corrēctum give us English ‘direct’, ‘erect’ and ‘correct’ respectively.

Germanic, meanwhile, provides us with a different set of reflexes of this verb. While we have already seen the rich set relating to wealth and kingship, the ‘straighten’ meaning of *h3reg- results in other interesting links. We have the (originally separate) verb and noun rake, a device for making straight lines, and the former participle right, originally meaning ‘straightened, directed’. Then we have reckon, perhaps a natural extension of the metaphor of lining things up in order to count them. Finally, from a causative ‘make straighten up’ we have reach (as if ‘straightening out one’s arm’).

This here is the greatest joy of etymology for me; by untangling these webs of relationships, we can show how so much of our vocabulary results from variations upon a common root. It reminds us of the continual creativity involved in using language and, by extension, the creativity of language users, i.e., humans.

Whisky Galore and A Go Go!

December 3, 2022 John Hutchinson Comments 0 Comment

When it comes to etymology, most words have a somewhat mundane route into a language: they either are retained from a direct ancestor or were borrowed at some point from another language. Within the latter category, these words tend to come in batches, often either through an intensive period of contact between peoples, as with the Old Norse loans into English, or through the importation of specific vocabulary which related to aspects of culture which were being borrowed from the group in question, such as e.g law terms deriving from the French used in English courts after the Norman Conquest.

However, every so often, there come along lexical items with a significantly more complex and idiosyncratic path into a language, and occasionally words may interplay with one another in interesting ways. We find such a complex interplay with galore and agogo.

Galore by itself is already an interesting form, as it is one of a small number of loanwords from Gaelic (likely specifically Scottish gu leòr) which does not have some kind of connection with Gaelic culture or geography. This expression can mean either ‘enough’ or ‘much, plenty’, and occurs in several constructions as a result. For instance, in Scottish Gaelic when asked ‘how are you?’, one might respond ceart gu leòr ‘all right, OK’, literally ‘right enough’.

This phrase, in a number of varying spellings such as gilore or gallore, appears to have begun to arrive in English in the mid 17^th Century (or at least this is the date of the earliest citation in the Oxford English Dictionary). When this form was borrowed into English it underwent semantic shift and narrowing, coming to specifically mean ‘in abundance, plenty’, losing the sense of ‘enough’. It seems to have been somewhat colloquial in use, not being particularly frequent in writing, and is disproportionately concentrated in Scottish works, including an attestation in the journals of Walter Scott.

This form comes to its greatest in prominence in English through its use in a Compton Mackenzie novel and later Ealing comedy titled Whisky Galore! Both the novel and film centre on a remote Scottish island, and the novel in particular makes use of Gaelic throughout, so the use of ‘galore’ fits in well with the setting.

This work in particular, however, had a more interesting impact than simple popularity. As with many best-selling works, it received translations into other languages, and in this case the French translation was titled Whisky à-Gogo, deriving likely from the Old French gogue ‘fun’. This title then was itself used as the name of a nightclub in Paris, the world’s first discothèque. The concept rapidly grew in popularity, with Whisky à-Gogo venues spreading across the globe, as far as Papeete in Tahiti (and Cardiff!), the most famous probably being the the Whisky a Go Go on Sunset Strip in Hollywood. (In the English-speaking world gogo got split into two, possibly on false analogy with the verb ‘go’.)

A film poster for the film 'Roadrunner a go-go' — But there’s only one Roadrunner…

From here on ‘a go go’ or just ‘go go’ became a by-word for everything hip and cool (or ‘groovy’) in the 1960s. Go-go dancers dance in go-go clubs, of course, but the meaning became more and more nebulous over time. In cinema, 1965 was a banner year, with Roadrunner a go-go up against Monster a go-go. This year also an unsuccessful attempt to extend this—word? phrase?—by analogy, with the notorious Batman parody Rat Pfink a Boo Boo. Nobody seems to have got this (not terribly good) joke, and on subsequent reissues the film was “corrected” to Rat Pfink and Boo Boo. (You’re reading this etymology here first. Even the director who came up with the title didn’t realize it, but we’re linguists, we know better.) But the shelf life of terms denoting popular trends is short, and anyone using it now probably means for it to lend antiquated flavour of the swinging 60s. Contrast with galore, which retains its more generic use and seems unlikely to drop out of common usage in the near future.

Who are the Germans?

October 19, 2022 John Hutchinson Comments 1 comment

You may be familiar with the fact that the Germans refer to themselves as Deutsch and their country as Deutschland, and we find this term also in most other Germanic languages, such as Dutch Duits or Swedish Tysk, as well as Italian Tedesco. However, there are many other names in other parts of Europe. The French and Spaniards call them Allemand/Alemán, as do the Welsh with Almaenaidd; the various Slavic languages share a different term again, seen in e.g. Polish Niemiec or Russian Nemets. In the Baltic the Lithuanians and Latvians have their own terms not seen anywhere else (Vokietis and Vācijis respectively), while in Finland and Estonia they call them Saksi. We could also add some assorted forms from smaller languages, such as Miksas from Old Prussian, an extinct sister language to Lithuanian and Latvian.

An aerial shot of the meeting of the Rhine and Mosel rivers at Koblenz — The Deutsches Eck, or ‘German corner’, in Koblenz

Now, it is not unusual for inhabitants of a country to refer to themselves and their country with a different form from that used by outsiders (when was the last time you called China Zhongguo or India Bharat?). What is particularly notable about the German case, however, is the diversity even among its immediate neighbours. Contrast e.g. France, where everyone uses some form of derivative of Latin Francia (after the Germanic tribe the Franks), though the Greeks still call it Gallia after the Roman province of Gaul. Similarly, most call Spain some form derived from Hispania and Italy one from Italia. So, this diversity in names for the Germans requires some explanation.

Whence this plethora of terms? A consideration of history leads us to our answer. Recall that the modern country of Germany is a relatively recent creation, only being officially united in the mid 19^th century by Otto von Bismarck. While there was a political entity that occupied the area in the form of the Holy Roman Empire it was only a relatively loose collection of small states, and prior to that the area was inhabited by a number of distinct Germanic-speaking peoples.

As a result, some of these names derive from the individual groups or tribes which lived in part of the area: so in the Western Romance and Brittonic Celtic languages the name of the Alemanni tribe was applied to the Germans as a whole. The same process occurred in the northeast with the Baltic Finns and the Saxons: not only were the Saxons the nearest group, but also, due to a combination of the Hanseatic League controlling trade through the Baltic and the anti-pagan crusading of the Teutonic Knights (another Deutsch-relative, see below), many Saxons came to settle in the Eastern Baltic, with some of their descendants still living in Estonia and Latvia today. Some small varieties show different groups again: some of the smaller Germanic varieties use a form derived from Prussian, after the state which ended up uniting the German peoples.

English takes a slightly different approach, deriving the term Germans from the Latin name of the region; Germania. This term included two Roman provinces covering much of modern-day Belgium, Switzerland, parts of eastern France and the Rhineland in modern Germany, as well as applying to the larger swathe of barbarian territories further east. Interestingly, several languages use this term to refer to Germany the country despite using a different term to refer to the Germans: Italian and Russian are the most notable examples.

We find a different source again with the Slavic Nemets terms. There is again some dispute in origin, but the general consensus is that it derives from a Slavic root *němъ meaning ‘mute’, itself of contested origin. The meaning likely was not ‘mute’ necessarily, but rather simply denoted that these groups were not Slavic-speaking. This puts in a similar group to the word ‘barbarian’ in fact, which derives from a Greek word meaning ‘those who go bar-bar/talk incomprehensibly’. Similar origins to do with ‘talking’ are likely behind the Baltic Vok-/Vāc-/Miks- forms as well.

Finally, what of German ‘Deutsch’? Well, as is the case with many endonyms it is a relatively simple and self-referential etymology. It ultimately derives from an Indo-European root *tewteh2 meaning simply ‘people’, which shows up also in e.g. Irish túath with the same meaning. This form may also be the source of Romance forms such as Spanish todo or French tout meaning ‘everyone/everything’. This root even survives in Slavic, in Russian giving the form čužoj, meaning ‘foreign, alien’. This ended up as Germanic *þeudō, which through an adjective formation *þiudiskaz meaning something like ‘of the people’ ultimately leads to the modern German form. This form also gives Latin Teutones, a likely Celtic or Germanic tribe which lived in the North German region and was encountered by the Romans early in their expansion northwards.

So, as with many other terms, such as the aubergine words which have been discussed here before, the differences between languages are reflective of a complex history. In this case the wide array of disparate terms of different etymologies reflects the complex history of the entity involved, specifically the absence of a country that even called itself ‘Germany’ until the modern era, as well as the extent to which different groups of ethnic Germans have moved about in Europe.

Careful who you climb a tree near: Respect and taboo in Vanuatu

July 20, 2022 Mike Franjieh Comments 1 comment

One humid afternoon, during breadfruit season in North Ambrym, my language teacher, Isaiah, and I were on the lookout for some ripe breadfruit to roast for lunch. Our path led past his nephew, George’s, house. Isaiah saw some ripe breadfruit in the tree next to where George was sitting on his veranda. Isaiah wanted to get the breadfruit, but said that because George was there, he couldn’t, and we would have to find some others instead. I asked if it was George’s breadfruit tree, and that’s why he didn’t want to take it when George was around. Isaiah said no; rather, the problem was if we went up the tree when George was underneath, then he would have to pay a small fine to George. Over a lunch of roasted and pounded breadfruit called wuwu, Isaiah explained further. It was to do with respect and taboo.

Respect in language takes many forms. There is the tu/vous distinction in French, where tu is the informal form of ‘you (singular)’ and is used with friends and those younger than you, whereas vous ‘you (plural)’ is formal and is used with those elder or senior than you and for people you don’t know. Similar distinctions are found with the German du/Sie. English doesn’t have a grammatical distinction in politeness like this, but uses different sentence structures to express politeness: compare pass me the salt please with could you please pass me the salt, or the even more polite would you be so kind as to pass me the salt please.

Now let’s get back to eating that heavy sticky coconut-cream-slathered wuwu with Isaiah. He told me that you must respect certain members of your extended family by showing physical politeness. Respect is translated as tengnean in the language of North Ambrym. The people who you must respect are your taboo family, described by the verb gorrne. Respect for your taboo family on Ambrym is realised in different ways – through physical restrictions and through language. The family members who command the most respect are your sister’s son or your husband’s brother.

The physical restrictions with a taboo relative include:

You can’t eat in front of them
You can’t joke with them
You can’t climb over them, or be physically higher than them
You can’t sleep in front of them
You can’t enter their house

But what about restrictions on language? The normal translation of ‘hello’ in North Ambrym would be neng le, which literally means ‘you there’, using neng, the singular form of ‘you’. But you are not allowed to say this to your taboo relatives. Instead, you must say gōmōro le using the dual form of ‘you’, meaning ‘you two there’, even though you are addressing one person. This is similar to French or German mentioned earlier. However, North Ambrym, like many Oceanic languages, not only has singular and dual, but also paucal, meaning ‘a few’, and plural pronouns. Of these possibilities, the dual is used for respect, not the plural as in French or German.

Respect is not confined to pronouns such as ‘you’; people also have to avoid using certain words in front of their taboo relatives. For example, if your sister’s son came, and you invited him to sit down and have some food, you would have to avoid certain verbs, such as taa ‘sit’ or ngene ‘eat’. You would use lingi ‘put’ instead of ‘sit’ and tewene ‘make’ instead of ‘eat’ so the whole sentence would be rephrased as ‘you-two come and put your-dual-self here and make the food’.

You must also avoid certain words concerning body parts, specifically words relating to parts of the head. Normally when talking about body parts in North Ambrym you would use a bound noun – a type of noun which specifies who owns the body part – so the word for ‘tooth’ would be lowo-n ‘his/her tooth’, lowo-m ‘your tooth’, or lowo-ng ‘my tooth’. The end of the noun (-n/-m/-ng in this example) indicates whose tooth it is. But these words are not allowed when talking in front of your taboo relatives. Instead, you could use a free form of the noun, such as leo ‘tooth’.

Another avoidance strategy is to change a verb to a noun using a special nominalising prefix a- that appears on the beginning of the word and turns it into a noun. The verb itself is also reduplicated. For example, the verb ta ‘cut’ can be turned into a noun atata ‘tooth’ (literally ‘thing for cutting’).

Finally, a more idiomatic expression could be used; in this case, tooth is replaced by tō which literally translates as ‘limpet shell (traditionally used as a vegetable grater)’ or teye ‘clam shell/axe’ as a way of avoiding the bound form for ‘tooth’.

Here’s a handy table to help you get your head (or just head!) around avoiding the bound forms.

Bound	Free	Nominalisation	Idiomatic
rralnye-n ‘his, her ear’	teleng ‘ear’	arorongta ‘thing for listening, headphones’	harrlengleng ‘listening’
lowon ‘his, her tooth’	leo ‘tooth’	atata ‘thing for cutting’	tō ‘limpet shell (used as a grater)’ teye ‘clam shell, axe’
metan ‘his, her eye’	marr ‘eye’	ateter ‘thing for seeing, glasses’	hal ‘road, path’ glas ‘glasses’
guhun ‘his, her nose’	kuu ‘nose’	akunuknuu ‘thing for smelling’
woulun ‘his, her hair’	wovyul ‘hair’		ōrr ge mre ‘place which is above’

As time passes, so do traditions, and the older generations mourn the loss of respecting their taboo relatives. They complain that younger generations now joke with their taboo relatives or put their arms around them. This art of speaking is being lost and the physical taboos are being eroded. However, this change is not new and has been going on for several generations. Some of the more extreme forms of respect are almost out of living memory. One of the village elders, Ephraim, recounted a memory of seeing how his grandmother, Mataran, displayed respect when returning from the garden, with her vegetables one day. When she approached her home, she saw that one of her husband’s brothers was there. She came close, then crawled the rest of the way past her husband’s brother with her basket of vegetables over her shoulder, until she was in her doorway before standing up again.

So the next time you are in Vanuatu, take care when climbing trees and make sure you know which of your relatives are nearby!

The Story of Aubergine

June 22, 2022 Steven Kaye Comments 2 comments

As the University of Surrey’s foremost (and indeed only) blog about languages and how they change, MORPH is enjoyed by literally dozens of avid readers from all over the world. But so far these multitudes have not received an answer to the one big linguistic question besetting modern society. Namely, what on earth is going on with the name of the plant that British English calls the aubergine, but that in other times and places has been called eggplant, melongene, brown-jolly, mad-apple, and so much more? Where do all these weird names come from?

I think the time has finally come to put everyone’s mind at rest. Aubergines may not seem particularly eggy, melonish, jolly or mad, but lots of the apparently diverse and whimsical terms for them used in English and other languages are actually connected – and in trying to understand how, we can get some insight about how vocabulary spreads and develops over time. It turns out that one powerful impulse behind language change is the fact that speakers like to ‘make sense’ of things that do not inherently make sense. What do I mean by that? Stay tuned to find out.

Long purple aubergine

To get one not-so-linguistic point out of the way first, there is no real mystery about eggplant (the word generally used in the US and some other English-speaking countries, dating back to the 18th century), which is not linked to anything else I am talking about here. It is hard to imagine mistaking the large, purple fruit in the photo above for any kind of egg, but that is not the only kind of aubergine in existence. There are cultivars with a much more oval shape, and even ones with white rather than purple skin: pictures like this, showing an imposter alongside some real eggs, make it obvious how the word eggplant was able to catch on.

Meanwhile, aubergine, which is borrowed from French as you might expect, has a much more complex history, and can be traced back over many centuries, hopping from language to language with minor adjustments along the way. The plant is not native to the US, Britain or France, but to southern or eastern Asia, and investigating the history of the word will eventually take us back in the right geographical direction. Aubergine got into French from the Catalan albergínia, whose first syllable gives us a clue as to where we should look next: as in many al- words in the Iberian peninsula (e.g. Spanish algodón ‘cotton’), it reflects the Arabic definite article. So, along with medieval Spanish alberengena, the Catalan item is from Arabic al-bādhinjān ‘the aubergine’, where only the bādhinjān bit will be relevant from here on. This connection makes sense, because the Arab conquest had such an impact on the history of Iberia. And more generally, we have the Arabs to thank for the spread of aubergine cultivation into the West, and also – indirectly – for this charming illustration in a 14th-century Latin translation of an Arabic health manual:

Illustration featuring three people in front of a stand of aubergine plants — Page from the 14th c. Tacuinum Sanitatis (Vienna), SN2644

But bādhinjān is not Arabic in origin either: it was borrowed into Arabic from its neighbour, Persian. In turn, Persian bādenjān is a borrowing from Sanskrit vātiṅgaṇa… and Sanskrit itself got this from some other language of India, probably belonging to the unrelated Dravidian family. The word for aubergine in Tamil, vaṟutuṇai, is an example of how the word developed inside Dravidian itself.

That is as far back as we are able to trace the word. But the journey has already been quite convoluted. To recap, a Dravidian item was borrowed into Sanskrit, from there into Persian, from there into Arabic, from there into Catalan, from there into French, and from there into English – and in the course of that process, it managed to go from something along the lines of vaṟutuṇai to the very different aubergine, although the individual changes were not drastic at any stage. The whole thing illustrates how developments in language can go with cultural change, in that words sometimes spread together with the things they refer to. In the same way, tea reached Europe via two routes originating in different Chinese dialect zones, and that is what gave rise to the split between ‘tea’-type and ‘chai’-type words in European languages:

[Map created by Wikimedia user Poulpy, licensed CC BY-SA 3.0, cropped for use here]

This still leaves a lot of aubergine words unaccounted for. But now that we have played the tape backwards all the way from aubergine back to something-like-vaṟutuṇai, we can run it forwards again, and see what different historical paths we could follow instead. For example, Arabic had an influence all over the Mediterranean, and so it is no surprise to see that about a thousand years ago, versions of bādhinjān start appearing in Greece as well as Iberia. Greek words could not begin with b- at the time, so what we see instead are things like matizanion and melintzana, and melitzana is the Greek for aubergine to this day. There is no good pronunciation-based reason for the Greek word to have ended up beginning with mel-, but what must have happened is that faced with this foreign string of sounds, speakers thought it would be sensible for it to sound more like melanos ‘dark, black’, to match its appearance. That is, they injected a bit of meaning into what was originally just an arbitrary label.

Meanwhile the word turns up in medieval Latin as melongena (giving the antiquated English melongene) and in Italian as melanzana, and a similar thing happened: here mel- has nothing to do with the dark colour of the fruit, but it did remind speakers of the word for ‘apple’, mela. We know this because melanzana was subsequently reinterpreted as the expression mela insana, ‘insane apple’. To produce this interpretation, it must have helped that the aubergine (like the equally suspicious tomato) belongs to the ‘deadly’ nightshade family, whose traditional European representatives are famously toxic. So, again, something that was originally just a word, with no deeper meaning inside, was reimagined so that it ‘made sense’. As a direct translation, English started calling the aubergine a mad-apple in the 1500s.

Parody of the "Keep Calm and Carry On" posters, reading "You don't have to be mad to work here but it helps" — Poster from a 16th c. aubergine factory

There are many more developments we could trace. For example, I have not talked at all about the branch of this aubergine ‘tree’ that entered the Ottoman Empire and from there spread widely across Europe and Asia. But instead I will return now to the Arab conquest of Iberia. This brought bādhinjān into Portuguese in the form beringela, and then when the Portuguese started making conquests of their own, versions of beringela appeared around the world. Notably, briñjal was borrowed into Gujarati and brinjal into Indian English, meaning that something-like-vaṟutuṇai ultimately came full circle, returning in this heavy disguise to its ancestral home of India. And to end on a particularly happy note, when the same form brinjal reached the Caribbean, English speakers there saw their own opportunity to ‘make sense’ of it – this time by adapting it into brown-jolly.

Brown-jolly is pretty close to the mark in terms of colour, and it is much better marketing than mela insana. But from the linguist’s point of view, they both reinforce a point which has often been made: speakers are always alive to the possibility that the expressions they use are not just arbitrary, but can be analysed, even if that means coming up with new meanings which were not originally there. To illustrate the power of ‘folk etymology’ of this kind, linguists traditionally turn to the word asparagus, reinterpreted in some varieties of English as sparrow-grass. But perhaps it is time for us to give the brown-jolly its moment in the sun.

SMG – I’d Arapaho, Roon, Sala, Tubar and Nara, but alas no Oroha paradigms

September 21, 2021 Greville Corbett Comments 3 comments

A palindrome is a linguistic delight: it reads the same in both directions. For example: level. Or Anna, or indeed Hannah. This is a visual trick: if you record yourself saying one of these words and play the recording backwards, it won’t sound exactly the same.

Palindromes hit the big time in the parrot sketch. They were also promoted by ABBA, with their top hit SOS!

Here’s a nice one from North Ambrym (an Oceanic language spoken in Vanuatu): rrirrirr ‘sound a rat makes when you try and kill it but you miss it’. And a long one from Estonian: kuulilennuteetunneliluuk ‘bullet flying trajectory tunnel’s hatch’. I’m not sure that one is used much (except in blogs about palindromes).

We can go up a level (!), as it were, to palindromic phrases. A famous one of these is:

A man, a plan, a canal – Panama!

This has been around at least since 1948. It has often been extended, as in this version due to Guy Jacobson:

A man, a plan, a cat, a ham, a yak, a yam, a hat, a canal – Panama!

And here’s a Russian sentence palindrome: Рислинг сгнил, сир. ‘the Riesling has gone off, sir’ More Russian palindromes at https://bit.ly/3AtxBID. For French sentence palindromes go to https://bit.ly/3kmC5LE. And there are even songs based on such palindromes:

They have palindromes in American Sign Language:

Not surprisingly, palindromes don’t translate. Though we can go up another level (!) of cleverness, to the bilingual palindrome: I love / e voli. This is half English, half Italian, and overall a palindrome. More of these at https://bit.ly/39ohoZy. It’s truly amazing what people can create, including whole poems as palindromes: https://bit.ly/3tTWtaa.

Some time ago, I mentioned to linguist colleagues that Malayalam (a Dravidian language of southern India) is a palindromic language. One colleague’s eyes opened wide, and he asked whether it was palindromic at the word level or the sentence level. What a great idea! Of course, it’s just the name which is a palindrome (just as Anna is a palindrome but that doesn’t make Anna a palindromic person – there are deep issues here: what does a name refer to?).

It turns out that there are over seventy “palindromic languages”, including some that are central to our research in SMG, notably Iaai (spoken in New Caledonia). Here are some more: Efe, Ewe, and Atta.

What then of E (also called Wuse/Wusehua), a Tai-Chinese mixed language, of Guangxi, China? Yes, it’s a palindrome, just not a very impressive one. Just as the English pronoun I is a palindrome, though hardly one to get excited about (unless you’re called Anna or Hannah of course). But it gets much better. You may have noticed that linguists increasingly give three letter codes after language names. These are the ISO codes that we use to uniquely identify a language, to make sure that we’re talking about, say, the language Aja (a Nilo-Saharan language of Sudan), ISO code aja, and not Aja (a Niger-Congo language of Benin), ISO code ijg. So, what is the ISO code for the language E? It’s eee. The language name and the code are both palindromes! Similarly there’s U (an Austroasiatic language of the Yunnan Province of China), ISO code uuu.

Here are the languages which are doubly palindromic (name and ISO code):

Name	ISO code
E	eee
Efe	efe
Ewe	ewe
Iaai	iai
Kerek	krk
Naman	lzl
Mam	mam
Nen	nqn
Ofo	ofo
Ososo	oso
Utu	utu
U	uuu
Yoy	yoy

A real star is Naman, whose ISO code is quite different, lzl, but still palindromic. Where does that come from? Well, the language has an alternative name, Litzlitz, so when it’s not a palindrome it’s a reduplication!

Back to the tricky use of “palindromic language”. Iaai is a palindromic name. As we’ve seen, its ISO code iai is also a palindrome. And the language does have some very nice palindromes:

aba ‘caress’
ee ‘locative – near the interlocuter’
ii ‘to suck’
iei ‘to hurt, cause pain’
ikiiki ‘repugnant’
iwi ‘rudder’
komok ‘sick’
maam ‘your manner’
mem ‘Napolean fish (Cheilinus undulatus)’
omoomo ‘women’
nokon ‘his/her infant’
oṇo ‘Barracuda (Sphyraena sp.)’
öö ‘spear’
ölö ‘mount, embark, disembark’
ölö ‘legume (Pueraria sp.)’
u ‘an old word for yam’
uu ‘fall from a height, chop down (of tree)’
ûû ‘a dispute, to dispute’
ûcû ‘similar, same’ (a nice meaning for a palindrome!)
ûcû ‘to exchange, buy, shop’

It would be impressive if you could read this post backwards, and have it make sense. But that wouldn’t be a BLOG but a GLOB, the latter being is an instance of a Semordnilap, but that is another story. For now, we welcome your favourite palindromes, in any language, in the comments.

For examples, thanks to Jenny Audring, Sacha Beniamine, Marina Chumakina, Mike Franjieh, Erich Round and Anna Thornton, and for the title (you’ve guessed what sort of title that is!), thanks to Steven Kaye.

Poolish

April 24, 2019 Matthew Baerman Comments 3 comments

Those who have out of desire have chosen to or out of dire necessity been forced to bake their own bread may have encountered the term poolish. It refers to a semi-liquid pre-ferment used in bread-making, a mixture of half water and half white flour mixed with a teeny bit of yeast and allowed to slowly ferment for several hours, up to a day, before mixing up the final dough.

The word itself is an exceedingly odd one, and has been the source of much head-scratching and inconclusive speculation among bread-bakers across the world: it looks like the English word Polish, but is spelled funny, and anyway seems to be borrowed from French, where the spelling would be funnier still. Most discussions of the technique include the obligatory etymological digression, usually fantastical, involving journeymen Polish bakers fanning out over Europe. Linguists too have gotten on the trail: David Gold’s Studies in Etymology and Etiology (2009) devotes a whole page to the question, but does not get too far.

In its current form it is technical jargon from French commercial baking, and has probably made its way to a broader public through Raymond Calvel’s influential Le gout du pain (‘The taste of bread’) from 1990. In his account:

This method of breadmaking was first developed in Poland during the 1840s, from whence its name. It was then used in Vienna by Viennese bakers, and it was during this same period that it became known in France. (2001 edition translated by Ronald Wirtz)

This explanation has been widely accepted, and appears in one form or another in any number of bread-baking books. But how could it even be true? The first problem is the word itself. Poolish is not the French word for Polish, and doesn’t much look a French word anyway. In earlier French texts it crops as pouliche, which looks more French and is indeed the word for a young mare, whose connection to bread dough is tenuous at best. But earlier French texts also have the spelling poolisch or polisch, which looks rather more German than French and suggests we follow the Viennese trail instead.

This thread of inquiry has its own potential hiccoughs. The German word for Polish is polnisch, with an [n], so would this not just be fudging things? Actually not: polisch, poolisch, pohlisch or pollisch turn up often enough in older texts as alternative words for ‘Polish’, particularly in southern varieties of German that include Austria. And it is exactly in these form that we find it being used to refer to this particular process, juxtaposed with Dampfl (or Dampfel or Dampel), the term in southern Germany and Austria for a rather stiffer pre-ferment which goes through a shorter rising period, as in these two examples from 1865, one from Leopold Wimmer’s self-published advertising advertising screed for St. Marxer brand (of Vienna) pressed yeast, where it turns up as Pohlisch:

the other from Ignaz Reich’s (of Pest, as in Budapest) account of ancient Hebrew baking practices, where it’s rendered as pollisch.

The term polisch (in all its variants) in this sense seems to have died a natural death in German, only to reemerge during the current craft-baking revival in the guise of poolish.

But if poolish was originally the (or a) German word for Polish, we run up against the sticky question of what it was actually referring to. Calvel repeats the story that this technique was invented by Polish bakers (which turns up in a 1972 article in The Atlantic Monthly, I think anyway, because it’s but coyly revealed by Google in snippet view), a supposition which lacks as much plausibility as it does historical attestation. Poland has traditionally been a land of sourdough rye bread. Is seems unlikely that a novel technique involving the use both of white wheat flour and commercial pressed yeast (a relatively new product) would have been devised there and introduced into the imperial capital that was Vienna. So what on earth could it have meant?

Here I make my own foray into speculation; you read it here first. Poland is not just a land of sourdough rye bread, it is a land of a soup made from rye sourdough: żur or żurek (itself derived from sur, one variant of the German word for ‘sour’), still widely consumed and also sold in ready form form for time-strapped gourmands. Since the Austro-Hungarian Empire included much of what had once been Poland, it isn’t too far-fetched to think that people in Vienna might have been familiar with this soup. And since the salient characteristic of poolish is that it is basically liquid, in opposition to more solid doughs, my guess is that the term poolish arose as a facetious allusion to żur: a soup-like fermenting dough mixture, like the thinned-out sourdough soup that Poles eat.

This theory has the minor drawback of lacking any positive evidence in its favor. So far the only 19^th century reference to żur outside of its normal context that I have been able to find is as a cure for equine distemper, otherwise known as ‘strangles’. That leads us into the topic of pluralia tantum disease names…

Sense and polarity, or why meaning can drive language change

January 30, 2019 Jérémy Pasquereau Comments 0 Comment

Generally a sentence can be negative or positive depending on what one actually wants to express. Thus if I’m asked whether I think that John’s new hobby – say climbing – is a good idea, I can say It’s not a good idea; conversely, if I do think it is a good idea, I can remove the negation not to make the sentence positive and say It’s a good idea. Both sentences are perfectly acceptable in this context.

From such an example, we might therefore conclude that any sentence can be made positive by removing the relevant negative word – most often not – from the sentence. But if that is the case, why is the non-negative response I like it one bit not acceptable, odd when its negative counterpart I don’t like it one bit is perfectly acceptable and natural?

This contrast has to do with the expression one bit: notice that if it is removed, then both negative and positive responses are perfectly fine: I could respond I don’t like it or, if I do like it, I (do) like it.

It seems that there is something special about the phrase one bit: it wants to be in a negative sentence. But why? It turns out that this question is a very big puzzle, not only for English grammar but for the grammar of most (all?) languages. For instance in French, the expression bouger/lever le petit doigt `lift a finger’ must appear in a negative sentence. Thus if I know that John wanted to help with your house move and I ask you how it went, you could say Il n’a pas levé le petit doigt `lit. He didn’t lift the small finger’ if he didn’t help at all, but I could not say Il a levé le petit doigt lit. ‘He lifted the small finger’ even if he did help to some extent.

Expressions like lever le petit doigt `lift a finger’, one bit, care/give a damn, own a red cent are said to be polarity sensitive: they only really make sense if used in negative sentences. But this in itself is not the most interesting property.

What is much more interesting is why they have this property. There is a lot of research on this question in theoretical linguistics. The proposals are quite technical but they all start from the observation that most expressions that need to be in a negative context to be acceptable are expressions of minimal degrees and measures. For instance, a finger or le petit doigt `the small finger’ is the smallest body part one can lift to do something, a drop (in the expression I didn’t drink a drop of vodka yesterday) is the smallest observable quantity of vodka, etc.

Regine Eckardt, who has worked on this topic, formulates the following intuition: ‘speakers know that in the context of drinking, an event of drinking a drop can never occur on its own – even though a lot of drops usually will be consumed after a drinking of some larger quantity.’ (Eckardt 2006, p. 158). However the intuition goes, the occurrence of this expression in a negative sentence is acceptable because it denies the existence of events that consist of just drinking one drop.

What this means is that if Mary drank a small glass of vodka yesterday, although it is technically true to say She drank a drop of vodka (since the glass contains many drops) it would not be very informative, certainly not as informative as saying the equally true She drank a glass of vodka.

However imagine now that Mary didn’t drink any alcohol at all yesterday. In this context, I would be telling the truth if I said either one of the following sentences: Mary didn’t drink a glass of vodka or Mary didn’t drink a drop of vodka. But now it is much more informative to say the latter. To see this consider the following: saying Mary didn’t drink a glass of vodka could describe a situation in which Mary didn’t drink a glass of vodka yesterday but she still drank some vodka, maybe just a spoonful. If however I say Mary didn’t drink a drop of vodka then this can only describe a situation where Mary didn’t drink a glass or even a little bit of vodka. In other words, saying Mary didn’t drink a drop of vodka yesterday is more informative than saying Mary didn’t drink a glass of vodka yesterday because the former sentence describes a very precise situation whereas the latter is a lot less specific as to what it describes (i.e. it could be uttered in a situation in which Mary drank a spoonful of vodka or maybe a cocktail that contains 2ml of vodka, etc)

By using expressions of minimal degrees/measures in negative environments, the sentences become a lot more informative. This, it seems, is part of the reason why languages like English have changed such that these words are now only usable in negative sentences.

On prodigal loanwords

August 15, 2018 Jérémy Pasquereau Comments 0 Comment

Most people at some point in their life will have heard someone remark on how their language X (where X is any language) is getting corrupted by other languages and generally “losing its X-ness”. Today I would like to focus on one aspect of the so-called corruption of languages by other languages — lexical borrowings – and show that it’s perhaps not that bad.

European French (at least the French advertised by the Académie Française) is certainly a language about which its speakers worry, so much so that there is even an institution in charge of deciding what is French and what is not (see Helen’s earlier post). A number of English-looking/sounding words now commonly used in spoken French have indeed been taken from English, but English first took them from French!

For instance, the word flirter ‘to court someone’ is obviously adapted from English to flirt and it has the same meaning in both languages. But the English word is the adaptation of the French word fleurette in the expression conter fleurette! The expression conter fleurette is no longer used (casually) in spoken French.

“How could the universe live without your beauty?” “I wonder how sincere he is…”

Other examples of English words borrowed from (parts of) French expressions which then get adapted into French are in (2).

Thus un rosbif is an adaptation into French of roast beef which is itself an adaptation into English of the passive participle of the verb rostir “roast” which later became rôtir in Modern French, and buef “ox/beef” which later became boeuf in the Modern French.

The word un toast comes from English toast with the meaning “piece of toasted bread”. The English word itself was borrowed from tostée, an Old French noun derived from the verb toster which is not used in Modern French. The word pédigré comes from English pedigree but this word is itself adapted from French pied de grue “crane foot”, describing the shape of junctions in genealogical trees.

Pied de grue ‘Crane foot’

Finally, the verb distancer is transitive in Modern French, which means that it requires a direct object: thus the sentence in (a) is good because the verb distancer “distance” has a direct object, the phrase la voiture blanche “the white car”. By contrast, the construction in (b) is not acceptable (signified by the * symbol) because it lacks an object.

a. La voiture rouge a distancé la voiture blanche.
‘The red car distanced the white car.’
b. *La voiture rouge a distancé.

The (transitive) Modern French verb distancer comes from English to distance which itself is a borrowing from the no-longer-used Old French verb distancer which was uniquely intransitive with the meaning “be far” (that is, in Old French, distancer could only be used in a construction with no direct object).

Another instance is (3): the word tonnelle ‘bower, arbor’ was borrowed into English and became tunnel under the influence of the local pronunciation. The word tunnel was then borrowed by French to refer exclusively to …. wait for it … tunnels. Both words now subsist in French with different meanings.

Une tonnelle ‘a bower’, Un tunnel ‘a tunnel’

Other examples of words that were borrowed into English and ‘came back’ into French with a different meaning are in (4).

The ancestor of tennis is the jeu de paume during which players would say tenez “there you go” as they were about to serve (at that time the final “z” was pronounced [z], it is not in Modern French). This word was adapted into English and became tennis which was then borrowed back into French to refer to the sport jeu de paume evolved into.

Jeu de paume vs. tennis

The Middle French word magasin used to refer to a warehouse, a collection of things. This word was borrowed into English and came to refer to a collection of things on paper. The word magazine was then borrowed back into French with this new meaning.

The history of the word budget also interesting. The word bouge used to mean “bag” and a small bag was therefore bougette (the -ette suffix is used as a diminutive, e.g. fourche “pitchfork” – fourchette “fork”). The word was borrowed into English where its pronunciation was “nativized” and it came to refer to a small bag of money. It was then borrowed back into French with the new meaning of “allocated sum of money”. Finally, ticket was borrowed from English which borrowed it from French estiquet, which referred to a piece of paper where someone’s name was written.

This happens in other languages of course. For instance, Turkish took the word pistakion ‘pistachio’ from (Ancient) Greek which became fistik. (Modern) Greek then borrowed this word back from Turkish which was then spelled phistiki with the meaning ‘pistachio’.

The main lesson I draw from the existence of ‘prodigal loanwords’ is that one’s impressions of language corruption often lack the perspective to actually ground that impression in reality. A French speaker looking at flirter ‘flirt’ may think that this is another sign of the influence of English — and they would be right — without being aware that this is after all a French word fleurette just coming back home.

Do you know other examples of prodigal loanwords? Please, share by commenting on this post!

Sources:
L’aventure des langues en Occident, Henriette Walter
Honni soit qui mal y pense, Henriette Walter
Jérôme Serme. 1998. Un exemple de résistance à l’innovation lexicale: les “archaïsmes” du français régional, Thèse Lyon II
Javier Herráez Pindado. 2009. Les emprunts aller-retour entre le français et l’anglais dans le sport. Universidad Politécnica de Madrid.

MORPH

A blog about languages and how they change

Browsed by
Category: French