Digital resources in the Social Sciences and Humanities OpenEdition Our platforms OpenEdition Books OpenEdition Journals Hypotheses Calenda Libraries OpenEdition Freemium Follow us

Lockdown etymologies: chowder

The American English word chowder refers to a thick fish soup made in New England. It is said (here) to go back to late Latin caldaria ‘cooking pot’. More specifically: Fr. chaudrée, a thick fish soup of the Charentes region (recipe here).

Lockdown etymologies: dog

As known already to Albert Dauzat, a major source of names of adult domesticated animals is in the name of the young. Thus French poulet ‘chicken’, earlier ‘chick’; English chicken, earlier ‘chick’; French cochon 1268-71 ‘young pig’ > 1611 ‘adult pig’; English pig , earlier > ‘young pig’;  Chinese zhu1 ‘pig’, earlier ‘young pig’; Modern Greek σκυλί ‘dog’, < Classical Greek σκύλαξ ‘puppy’; Modern Chinese gou3 ‘dog’, earlier (c. 200 BCE) ‘hairless puppy’; etc.

This might be true of the English word dog too. The early meaning of ‘cub’ perhaps survives in the term dogies, referring to calves in the cowboy song Git along, little dogies (here).

The etymology of PRT *cáni ‘one’

(view clickable tree)

(up one level to Rukai-Tsouic master post)

Wolff (2010, sub *táni) noticed the connection between PRT *cáni ‘one’, Itbayaten tanih ‘alone’, Ratahan tani ‘to separate’ and Bare’e tani ‘independent’; add Kavalan tani, utani ‘some, several, a few’. Bril (p.c. April 20, 2020) finds tani ‘only’ in Natauran Amis. The Rukai-Tsouic innovation consists of the semantic shift  of *cáni  to ‘one’ out of an original meaning something like ‘alone, only’.

Reference

Wolff, John U. (2010) Proto-Austronesian phonology with glossary. 2 vols. Ithaca: Cornell Southeast Asia program publications.

PMP *inum ‘to drink’

(to clickable tree)

In this post  I argued that Proto-Eastern Walu-Siwaish (PEWS) innovated *nanum ‘water’, which competed with  PAN *daNum. PEWS *nanum is reflected in all the EWS languages of Taiwan save Paiwan, as well as in all the Kra-Dai branches. A direct reflex is however not to be found in the Malayo-Poynesian languages, where *daNum reigns as ‘water’. In PEWS there existed a denominal verb  *mi-nanum  ‘to get water, to drink’: *mi-N constructions meaning ‘to acquire, get, obtain, collect, N’ are common Actor-Focus verbs in EWS languages. This *mi-nanum is the probable source of the main MP word for ‘to drink’: *inum, *um-inum.

Shortly before PMP, AF *mi-nanum ‘to drink’ underwent syncope of the unstressed middle vowel, much as in pre-MP *paŋudaN ‘pandanus’, giving Tagalog pandán, Malagasy fandrana, Malay pandan, Old Javanese paṇḍan, etc. This gave *mi-nnum. As the noun part of this *mi-N construction was not recognizable anymore, *mi-nnum was reanalyzed as*m-innum, with prefixed m- allomorph of <um> in words beginning in vowels. The internal cluster was then regularized, giving PMP *m-inum ‘to drink’.

My one and only

In the Formosan data recorded by Ino (1998:25), Taokas and Papora, two west-coast Formosan languages, have for ‘one’ the following forms: Sinkon (Taokas) tanu, Hajyovan and Vudol (both Papora) tanu. The form occurs reduplicated in Tanatanahu and Hameyan (both Taokas): tata:anu, and as ta’a:nu in Varaval (Taokas). The base form may reconstruct as *tanu or *sanu (not *Canu, since Sinkon Taokas reflects *C as s: masa ‘eye’, sareina ‘ear’, samani ‘to cry’; and the same is true of Papora: masa, sarina, samani: ). This is etymon #30 in the ABVD. The PAN word for ‘one’ is generally acknowledged to be *isa, which  seems to favor *sanu over *tanu.

However  Natauran Amis tanu ‘only, just’ (Bril, p.c., 31 March 2020) can only go back to *tanu.  I hypothesize that the original meaning of *tanu was ‘only’, as in Amis, and that  in Papora and Taokas the word shifted to ‘only one’ and ultimately to ‘one’, displacing earlier forms containing *isa (contra Tsuchida 1982:10).

(updated Dec 27, 2022)

References

Ino, Yoshinori 伊能嘉矩. 1998. 巡臺日程 Xún Tái Rì Chéng [A Formosan itinerary]. In 伊能嘉矩番語調查手冊 Ino Yoshinori fānyǔ diàochá shǒucè [Ino Yoshinori’s field records], ed. by : T. Moriguchi, pp. 13-201. Taipei: Southern Materials Center.

Tsuchida, Shigeru. 1982. A comparative vocabulary of Austronesian languages of Sinicized ethnic groups in Taiwan. Part I: West Taiwan. Memoirs of the Faculty of Letters, University of Tokyo No 7.

Yet another word for ‘ten’ in Formosan languages

up one level to Eastern Walu-Siwaish

A hitherto undescribed etymon for ‘ten’ occurs in languages of the north and east coasts of Taiwan: Sakizaya, Puyuma, and north Formosan (Kavalan, Basay, Ketagalan).

In Sakizaya, Amis’s close relative, bataʔan serves for ‘ten’ in and multiples of ten (data collected by McNaught, 2017; see the Sakizaya page in Eugene Chan’s ‘numeral systems of the world’s languages’ (here). There no trace of a reflex of *puluq, unlike in Amis and in Puyuma.

Cauquelin’s Puyuma dictionary (Cauquelin 2015) only has puɭuʔ and məəp for ‘ten’ but ‘twenty’ is maka-bəʈaʔan. The prefix maka- serves in ‘ten’ and multiples of 10, eg maka-telun ‘30’, maka-pətəl ‘30’ etc. It is possible that maka-bəʈaʔan was simplified from maka-ɖua-bəʈaʔan: removing ɖua would cause no ambiguity.

Kavalan has Rabtin ‘10’, zusabtin ‘20’ where –btin is the numeral ‘10’ (Li & Tsuchida 2006). One source of Kavalan /i/ is *qa. Vowel syncope is common in Kavalan although the conditions governing this process have not been elucidated. So -btin < *b()t()qan. Kavalan merges *t and *C into *t, so *b()t()qan may originate in *baCaqan. Basay is similar to Kavalan: labatan ‘10’, lusa batan ‘20’, as is Ketagalan ɭabat-an ‘10’, ɭusa batan ‘20’ (Ogawa, cited in Ferrell 1969). All these forms are regular outcomes of *baCaq-an. Puyuma provides the evidence for reconstructing -C- as against -t-.The formative Kavalan Ra-, Basay la-, Ketagalan ɭa– is a distinct morpheme.

A tentative etymology can be offered. Tsuchida (1988) glossed the Bunun word bataqan (< baCaq-an, *bataq-an) in Qato and Idhokan dialects as ‘racks (L-shaped – to carry woods)’. Nihira’s Bunun vocabulary gives ‘a carrying board on the back’. The name of a carrying device for multiple objects is a potential source of ‘ten’. Bunun is a Walu-Siwaish language, like the languages where *baCaq-an occurs in ‘10’ or ‘20’: we may assign *baCaq-an ‘carrying device’ to Proto-Walu-Siwaish and the derivation out of it of a new word for ‘ten’ to Proto-Eastern-Walu-Siwaish.

References

Cauquelin, J. (2015). Nanwang Puyuma-English dictionary. Institute of Linguistics, Academia Sinica, Taipei, Taiwan.

Ferrell, R. (1969) Taiwan Aboriginal groups: problems in cultural and linguistic classification. Monograph No. 17, Institute of Ethnology, Academia Sinica. Nankang: Academia Sinica.

Li, P. J-K, and S. Tsuchida (2006) Kavalan dictionary. Language and Linguistics monograph series A19. Nankang: Institute of Linguistics (preparatory office), Academia Sinica.

Nihira, Y. (1983) A Bunun Vocabulary (2nd edition). Privately published.

Tsuchida, S. 1988. Comparative word lists of Bunun dialects. Report of the research carried out in 1983. Unpublished manuscript.

The etymology of *puluq ’10’, at last

(Jump to Puluqish master post)

In Sagart (2008), a modification of the phylogeny in Sagart (2004), I set up a Puluqish group of Austronesian languages including all of PMP, all of Kra-Dai, plus three southern Formosan languages: Amis, Puyuma and Paiwan. This group was defined by the innovation *puluq for ‘ten’ (another innovation has since been identified, see the master post for Puluqish).  My subgrouping is in contrast to the standard view (e.g., in Blust’s phylogeny) that *puluq existed in PAN. I was not able, at the time, to explain how *puluq arose. My inability to do so at the time was interpreted as an indication that *puluq was etymologically opaque, as befits a PAN form, just like the numerals 1-4. If,  on the contrary,  *puluq was an innovative form, as I claimed, it must originate in another word or expression shifting to ’10’ just before proto-Puluqish, and traces of that original expression should be found, hopefully in modern Puluqish languages, or not too far from them on my tree.

The etymology of *puluq ‘ten’ can now be given. Amis has a root /poloʔ/, orthographic polo’ < *puluq, which occurs in all Amis dictionaries. Below I discuss three forms in Namoh’s large (Central) Amis-Chinese dictionary (Namoh 2013), where the most detailed information can be found:

  • si-polo’ 因分離而獨居,或分家 ‘living alone or separated due to separation’
  • ma-sipolo’. 分離,分闊,分居,單身,未婚 ‘separate, split, separation, single, unmarried’. Example sentence: Milaliw ko fafahi ni Kuraw, saka masipolo’ ciira i matini. 古饒的妻子昨天離家出走,所以他現在是單身 ‘Kuraw’s wife left home yesterday, so he is now single’
  • ma-si-polo’-ay (deverbal noun out of ma-sipolo’). 寡婦,搞婦,穌夫。單身 ‘widow, widower; single person’.

The Amis affixes ma- (stative) and -ay (nominalizing) are well-known. masipolo’ is a stative verb based on sipolo’; masipolo’ay is a deverbal noun based on masipolo’. It is not entirely clear what the function of si- in sipolo’ is. There is a prefix si- in Amis with existential or possessive meaning ‘have, be’ (Bril, p.c., circa 2020) .  In the English index at the end of Namoh’s dictionary, masipolo’  is given the gloss  ‘separated from and left alone’.  One could gloss si-polo’ as ‘for a state of separation to exist’; ma-si-polo’ as ‘be in a state of separation’, and ma-si-polo’-ay as ‘a N in a state of separation’. A gloss like ‘to separate from, and leave alone’ can be supposed for the stem polo’.

It is easy to see how polo’ could be put to use in counting numbers ten and above: the meaning ‘separate from, and leave alone’ well describes the mental operation of isolating n sets of ten objects and putting them aside before expressing the remainder.  For instance, ’35’ can be expressed as ‘three sets of ten, put them aside, and five’.

The process by which *N-puluq became the main form for ’10’ in the Puluqish languages can be reconstructed thus. The word for ‘10’ in Proto-Eastern-Walu-Siwaish (the immediate ancestor pf Proto-Puluqish) was *baCaqan (here). ‘35’ in that language might have been expressed simply as telubaCaqan-lima ‘3 times ten, five’, or more explicitly as telubaCaqanpuluq-lima ‘3 times 10, leave those alone, five’. *baCaqan being redundant, *N-baCaqanpuluq-N was simplified to N-puluq-N in Proto-Puluqish, with *N-puluq acquiring the meaning ‘N times 10’. But *baCaqan was predictably retained in northerrn Puluqish languages where there was no remainder to ‘leave alone’, i.e. in ’10’ and ’20’:

  • Sakizaya t͡sət͡saj a bataʔan ‘10’ (one-baCaqan), tusa bataʔan ‘20’ (two-baCaqan),
  • Puyuma maka-bʈaʔan ‘20’

We have a small but independent piece of evidence confirming  that *puluq and its derivatives were part of the vocabulary of counting: Pokpok Amis nom-a-sipolo’ ‘7’ (Li and Toyoshima 2006, #83l.1). This is a subtractive numeral: ‘three left aside’, where nom reflects *Nem ‘3’ (details here).  Here sipolo‘ does not directly refer to ‘ten’ but to the act of ‘leaving (three units, out of ten) aside’.

A consequence of the above is that in proto-Puluqish, ‘ten’ must have been not *puluq, but *sa-puluq with *sa-, the short form of the PAN numeral *isa/*esa ‘one’, preceding *puluq: ‘one (set of ten), put aside’.  Paiwan (Sagaran dialect, recorded by Hsiou-Chuan Chang) tapuluʔ  ‘ten’ directly reflects proto-Puluqish *sa-puluq. *sa-puluq is also the form that can be reconstructed for PMP (Blust’s Austronesian comparative dictionary).

In Kra-Dai, Ostapirat (2000) reconstructs Proto-Kra *pwlot, and Norquest (2015) gives Proto-Hlai *fu:t. Ostapirat (2004) had Proto-Hlai *apu:c, with final *c  (an error in my opinion; see here).  The vowel *a before the string *pu:c appears to reflect the *a in proto-Puluqish *sa-. As reconstructed by Ostapirat, proto-Hlai occasionally retains the first vowel of Austronesian words, but never the consonant before that vowel: proto-Hlai *ata A < *maCa ‘eye’, aka:i C < *Caqi[] ‘excrement’, *ura:ŋ A < *qudaŋ ‘shrimp’,  utu A ‘head louse’ < *quCuH2, *ipan < *nipen or *lipen ‘tooth’ etc.

Thus proto-Puluqish *sa-puluq is an adequate source of the Paiwan, proto-Malayo-Polynesian and proto-Hlai words for ‘ten’. Lack of *sa- in Amis and Puyuma puɭuʔ and polo’ is due to a simplifying innovation: there was no need for *sa- after the meaning ‘ten’ became entrenched with *puluq and the original semantics were lost. This innovation should be added to the six shared innovations of Amis and Puyuma (a.k.a ‘northern Puluqish’) described here.

Although Puyuma shows no trace of the *sa- in *sa-puluq, indirect evidence that it once possessed the long form exists: the serial counting word for ‘nine’ in Nanwang Puyuma, siwa, can only be explained as the result of analogical alignment on the *s- at the beginning of *sa-puluq (more on this here).

I am not aware of cognates of polo’ with semantics related to ‘separated, put aside’ outside of Amis. It is perhaps significant that the lexical source of the Puluqish word for ‘ten’ comes from Amis, a Puluqish language.  If *puluq were a PAN word, that would be a coincidence.

References

Blust, R. 2001. Malayo-Polynesian: new stones in the wall. Oceanic Linguistics 40, 1: 151-155.

Li, Jen-kuei and Masayuki Toyoshima (eds). 2006. comparative vocabulary of Formosan languages and dialects, by Naoyoshi Ogawa. Asian and African lexicon series 49. Institute for Languages and cultures of Asia and Africa, Tokyo University of Foreign Studies.

Namoh, Rata. 2013. O Pidafo’an to Sowal Misanopangcah [dictionary of the Amis language]. Taipei: Nan t’ien.

Norquest, Peter. 2015. A Phonological Reconstruction of Proto-Hlai. Brill.

Ostapirat, Weera. 2000. Proto-Kra. Linguistics of the Tibeto-Burman Area 23.1:1-251.

Ostapirat, Weera. 2004.  Proto-Hlai sound system and lexicons. in: Ying-ching Lin, Fang-min Hsu, Chun-chih Lee, Jackson T.S. Sun, Hsiu-fang Yang and Dah-an Ho (eds), Studies in Sino-Tibetan languages, papers in honor of Professor Hwang-cherng Gong on his seventieth birthday, 121-175. Nankang: Institute of Linguistics, Academia Sinica.

Sagart, L. (2004) The higher phylogeny of Austronesian and the position of Tai-Kadai. Oceanic Linguistics 43,2: 411-444.

Sagart, L. (2008) The expansion of setaria farmers in East Asia: a linguistic and archaeological model. In Sanchez-Mazas A, Blench R, Ross M, Peiros I, Lin M, eds. Past human migrations in East Asia: matching archaeology, linguistics and genetics, 133-157. Routledge Studies in the Early History of Asia, London: Routledge.

Faulty etymologies in the STEDT II: ‘to love’

One of the weaker Sino-Tibetan comparisons in the STEDT can be found here:

#1160 PTB *ŋ-(w)aːy COPULATE / MAKE LOVE / LOVE / GENTLE

The reconstructed form has detachable *ŋ- pre-initial (not a prefix), and a *w- onset marked as optional by means of parentheses.1 The possibility that forms with and without reflexes of these elements are unrelated is not given any consideration. In Matisoff’s Schrödinger-like version of comparative reconstruction, phonemes can be there, and not there.

At Sino-Tibetan level, ‘PTB *ŋ-(w)a:y’ is compared with Chinese 愛 ‘to love’, Middle Chinese  ‘ojH > modern ài.  The Chinese comparandum is vaunted as ‘excellent’ at the top of the STEDT page for this set.

愛 ài expresses a kind of feeling associated with the Confucian virtue of 仁 *niŋ > nyin > rén ‘kind(-ness), benevolence’.  Although its modern meaning is ‘to love’, and it is often translated by the verb ‘to love’, its more specific meaning in classical texts is  ‘to care for’,  as in this famous passage from the Confucian Analects 3.17/6/6, where in response to his disciple Zigong’s objection to the ritual sacrifice of a sheep, Confucius is quoted as saying “爾愛其羊,我愛其禮” “You love (=care for) the sheep, I love (=care  for) the rite”. Such is probably the original meaning.  The OC word for ‘to love’ was 字 *mə-dzə(ʔ)-s, correctly compared by Benedict to his PTB *m-dza ‘to love’.  In all Chinese dialects except (to my knowledge) the very archaic Waxiang and Caijia dialects, this word has been displaced by 愛 ài. ‘Love’ semantics, in the sense of a warm feeling one experiences for an individual person, are not old with 愛 ài.

A  minimal version of the comparison in STEDT had appeared in Benedict’s Conspectus (1972:192, fn.491): it related Proto-Karen *ʔai ‘love’ with 愛 ài. This would  be phonologically viable if the OC root ended in *-əj. In general, MC words with grave initials and rhyme -ojH (with ‘H’ for tone C, a.k.a qusheng) can go back to OC *-əj(ʔ)-s, *-ə(ʔ)-s, *-ək-s, *-ət-s or *-əp-s. Of these, only *-əj(ʔ)-s could potentially match forms in -a(:)j  outside of Sinitic. However, Baxter (1992) reconstructed *ʔɨt-s, with root-final *-t. Still, Zev J. Handel maintains at the bottom of the STEDT page for the item under review that *-j-s remains possible because OC rhyming does not provide direct evidence of contacts with *-t. While this is true, one should keep in mind that 愛 ài only rhymes once in the Shi Jing, and 僾 ài  ‘to pant, lose the breath’, whose phonetic is 愛 ài, also once: the opportunities for an unsuffixed final stop to surface in rhyming are thus quite limited. Absence of *-t (or *-p, for that matter) among the rhyme contacts of 愛 ài or 僾 ài  is not in itself evidence for a root ending *-j.

Word-families provides strong evidence for excluding *-j-s in 僾 ài ‘to pant, lose breath’, just cited: this word is the s-suffixed derivative of 唈 *qˤ[ə]p > ‘op > yì ‘short of breath’.  This is confirmed by the fact that  僾 ài  ‘to pant, lose the breath’ rhymes with 逮 modern dài < OC *m-rˤəp-s  ‘reach to’  in Ode 257. 逮 dài itself is the s-suffixed derivative of 眔 *m-rˤəp > dop > tà ‘reach to’.  There can be no question that 僾 ài and 逮 dài ended in *-p-s.

Baxter and Sagart (2014) accordingly reconstructed 僾 ài  as *qˤəp-s. This in turn cannot be reconciled with the view that 愛 ài ended in *-j-s: how could a word ending in *-j-s  be chosen as a phonetic for another ending in *-p-s ? evidently the root in 愛 ài  also ended in a stop. But which ?

While 愛 ài  ‘to care for’ was reconstructed with *-t-s in Baxter (1992), the OC contrast between *-p-s and *-t-s was lost at a late stage of OC: if the rhyming of 愛 ài   and 謂 *[ɢ]ʷə[t]-s > hjw+jH > wèi ‘say, tell, call’ in Ode 228 is based on a pronunciation in which the merger had taken place, 愛 ài may have had *-p-s earlier on, like 僾 ài. Accordingly Baxter and Sagart (2014) reconstructed 愛 ài as *[q]ˤə[p]-s, with square brackets around -[p] to convey the uncertainty between *-p and *-t.

Shortly after the publication of their book, and independently from it, Norquest (2015) arrived at the Proto-Hlai reconstruction *ʔə:p ‘to love’. He did not notice any connection to Chinese 愛 ài . Norquest’s *ʔə:p is not found in the other Kra-Dai branches,  in Austronesian or Austroasiatic. It is probably a Chinese loanword: this leaves no doubt that the Chinese word’s root ending was *-p.

Chinese root-final *-p is not reflected anywhere in the rest of Matisoff’s set: the likelihood that the Chinese form is related to the rest is nonexistent. The superficial resemblance between the Chinese and Karen forms is the fruit of convergence.

references

Baxter, W. H.  1992. A Handbook of Old Chinese phonology. Trends in Linguistics Studies and Monographs 64. Berlin: Mouton de Gruyter.

Baxter, William H. and Laurent Sagart. 2014. Old Chinese: a new reconstruction. New York: Oxford University Press.

Benedict, P.K. 1972. Sino-Tibetan: a Conspectus. Cambridge: University Printing House.

Norquest, Peter. 2015. A Phonological Reconstruction of Proto-Hlai. Brill.

  1. There is nothing wrong in itself with parentheses. Baxter and Sagart (2014) use parentheses in their OC reconstruction to signal phonemes whose presence or absence in a protoform are both compatible with all the comparative evidence at hand: for instance they reconstruct medial-r- between parentheses in 夫 *p(r)a > pju > fū ‘man’  because given Middle Chinese pju it is impossible to decide whether Old Chinese had -r- or not. Matisoff’s parenthesized phonemes are different: they are used to make manifest Matisoff’s decision to overlook a phoneme in making cognate decisions. []

Conflated cognate sets in the STEDT I — ‘excrement’

I have elsewhere (Sagart 2006) pointed out that sound correspondences are a minor part of the grounds on which the cognate sets in Matisoff’s 2003 book were assembled, despite Matisoff’s anger (Matisoff 2007) when the point was made in print. The same applies to the cognate sets in STEDT, which are a more evolved state of those in the book. Granted, to some extent relying on educated guesses is unavoidable in cognatising a large number of related languages: however, the sound correspondences in Matisoff’s book, which are his most recent statement on ST phonological history, are not sufficient to distinguish between phonetically and semantically resemblant etyma in ST.

Consider the STEDT cognate set #572 *kləy ‘excrement’ (here). In several of the languages where this putative etymon is reflected, doublets appear. The most prominent type includes one form with a liquid in its onset (column I) and another without (column II). See table 1.

 III
Kanauri (Sharma)s kli ‘urine’khə ‘shit’
Central Tsangla
(Egli-Roduner)
le ‘intestines’khi chung ‘buttock’
Bodo (Bhat)bi klə́ ‘liver’bi kí ‘excrement of fish’
Mikir (Grüssner)mék-krí ‘tears’hī ‘feces / shit’
Kayah
(Luangthongkum)
khrə¹¹ ‘body dirt’ci¹¹  ‘body dirt’
Newar (Genetti) ʈi (< kr-) ‘(ear)wax’khi  ‘shit’
Sunwar (Michailovsky)khriː ‘feces’kiː  ‘intestines’

Table 1: doublets in the STEDT cognate set #572 (accessed mid-may, 2019)

The differences in onsets and rhymes between the forms in the two columns are not explained anywhere in Matisoff (2003) or STEDT. They are not due to identified morphological processes applying on a single lexical root. We are dealing with two etyma: the first has an onset cluster, the second does not. Semantically the forms in the first column have more associations with ‘body dirt’ and those in the second column with ‘excrement’.

Taraon klɑi53, Proto-Northern Naga *C̥-kləy (French 1983), Written Tibetan lci < hlyi, Lepcha tə kli probably belong to the first etymon. Japhug Gyarong tɯ qe ‘excrement’ (Jacques 2015) probably belongs to the second: the Japhug onset cannot originate in a Cl-type cluster (Jacques, p.c.).

With the Chinese word 屎 MC syijX ‘excrement’, one source of MC sy- is OC *l̥-: if OC *l̥- itself can remount to PST *kl-, one seems to have a match to #kləy. However the Proto-Min initial for 屎 is *š-, which excludes OC *l̥-. An OC *l̥- would evolve to Proto-Min *tšh- (Baxter and Sagart 2014:93). Consequently, the Chinese word probably does not belong with the first etymon either. Baxter and Sagart reconstruct 屎 tentatively with a uvular initial, OC *[qʰ]ijʔ. None of the other sources of MC sy-: OC *s.t-, *n̥-, *ŋ̊-, *l̥-, are plausible in this word. Thus 屎 OC *[qʰ]ijʔ probably belongs to the second etymon, like Japhug Gyarong tɯ qe.

The first etymon, with the lateral cluster and ‘body dirt’ semantics, is best compared with Chinese 尸 *l̥əj ‘corpse’, via the notion of ‘carrion’.

The second etymon can be compared with Proto-Austronesian (Blust) *Caqi ‘excrement’. There are reasons within Austronesian to think that this word ended in a laryngeal or back-of-the-mouth fricative (here), although I am not sure anymore of the identity of that phoneme.

In my next post, I will discuss the STEDT *etymon #601 *m/s-tuːk ‘to spit’.

References

Baxter, William H. and Laurent Sagart. 2014. Old Chinese: a new reconstruction. New York: Oxford University Press.

French, Walter Thomas. 1983. Northern Naga: A Tibeto-Burman mesolanguage. New York, Univ., Diss. Ann Arbor : University Microfilms

Jacques, Guillaume. 2015. Dictionnaire Japhug-chinois-français 嘉绒-汉-法词典 Version 1.0.

Matisoff, J. A. 2003. Handbook of Proto-Tibeto-Burman. Berkeley, Los Angeles, London: University of California Press.

Matisoff, J. A. 2007. Response to Laurent Sagart’s review of Handbook of Proto-Tibeto-Burman: System and philosophy of Sino-Tibetan reconstruction. Diachronica 24,2: 435–444.

Sagart, L. 2006. Review: James A. Matisoff (2003) Handbook of Proto-Tibeto-Burman. System and philosophy of Sino-Tibeto-Burman Reconstruction. Diachronica 23,1: 206-223.

Is 袁 yuán ‘long robe’ a ghost word?a response to Guillaume Jacques

In a comment on a post of Guillaume Jacques’s I cited the word 袁 yuán ‘long robe’ as being cognate with Tibetan གོན་ gon “clothing”. This drew a response from Guillaume, from which I extract these passages:

“l’étymologie avec 袁 est difficilement acceptable philologiquement: la glose que tu cites « long robe » est la traduction de celle du shuowen 长衣貌, mais ce mot supposé est sans attestation textuelle (无书证); les dictionnaires ne citent que les gloses de dictionnaires: http://www.guoxuedashi.com/kangxi/pic.php?f=gxhz&p=2055”

and:

“Je pense que l’on ne peut pas utiliser de mots dont l’existence même n’est pas assurée pour faire du comparatisme (en plus quand bien même il aurait existé, la glose X貌 suggère qu’il devait plutôt s’agir d’un idéophone, pas d’un nom).”

Real words do not always occur in the received literature; for instance the Cantonese and Hakka popular word for “egg” has no early Chinese text occurrences, yet it corresponds well to the word “egg” in other ST languages (Baxter and Sagart 2014:324).

Specifically with regard to 袁:

The same word appears written as 褑 and 褤 in the Ji Yun 集韻 with the spelling 于元切= hjwon,and the gloss 衣也 “clothing”. The Ji Yun here is not citing the Shuo Wen, since the characters and the gloss wording are different.

The Shuo Wen itself which says “袁, 长衣貌也” is not citing an earlier dictionary/word list either, since there is no earlier attestation of the character, and the Shuo Wen is the first Chinese dictionary.

Dialectally the word is attested. Huang Kan 黄侃 in his《蕲春语》wrote that a long robe is called 長褑 in his own dialect, noting that the word must be the same as 褑 in the Ji Yun.

Although there are no early text examples of 袁, the character is attested both paleographically and as the head of a phonetic series. It includes 衣 “clothing” and a hand.

There is, then, no reason to assume that 袁 is not a real word. The comparison to WT gon is not problematic.

references

Baxter, William H. and Laurent Sagart. 2014. Old Chinese: a new reconstruction. New York: Oxford University Press.