Claims about the Egyptian genitive, checked against the whole TLA

corpus
TLA
genitive
Jansen-Winkeln and Werning both made claims about when Egyptian uses the direct and when the indirect genitive. What 16,812 attestations from the TLA make of them: one confirmed, one narrowed to a couple of formulas, one not there at all.
Published

16 September 2026

Egyptian has two ways of linking two nouns. The direct genitive puts them side by side, prw= nzw “the king’s house”; the indirect one inserts a connector, prw nï nzw “the house of the king”. For most of the language’s history both are available and both say the same thing. What decides the choice in a given instance has been argued about for a century.

Two pieces of work make claims specific enough to check against counts. Karl Jansen-Winkeln (ZÄS 127, 2000) goes through the older explanations — Sander-Hansen’s morphology, Junker’s ownership criterion, Grandet and Mathieu’s definiteness, Junge’s restriction, Schenkel’s nearer and wider belonging — and rejects all of them. What he puts in their place is that the choice is steered by the lemma and the form of the nomen regens, and he adds that working the rules out “would require extensive statistics”. Daniel Werning (2024, DOI 10.18452/28786) counts the genitives of the Book of Caverns, a composition of the 13th century BCE running to some 11,500 words: 526 attestations once the litany repetitions are set aside, 80 per cent of them direct against 20 indirect. He reads that as hypercorrection, overshooting not only the Late Egyptian of its own day but probably genuine Earlier Egyptian as well, and frames the whole of égyptien(s) de tradition as a mental translation from Late Egyptian as a first language into the emulated older stage as a second. Among his findings is that inanimate possessors prefer the direct genitive more strongly than animate ones — the reverse of what typology leads one to expect.

This post reports what comes out when the statistics are actually done. The material is the Thesaurus Linguae Aegyptiae, version 20, restricted to cases where both constructions were genuinely available: 16,812 attestations across the whole attested history.

The distribution across that history is uneven in kind as well as in quantity. From the New Kingdom onwards much of what survives is written in an older stage of the language rather than a contemporary one: 3,538 attestations are labelled égyptien(s) de tradition and a further 1,777 are Middle Egyptian written after the Middle Kingdom, against 906 in Late Egyptian of any sort. Almost half the corpus carries no language label at all, so those are lower bounds.

The configurations in which the direct genitive is structurally out of reach are listed by Jansen-Winkeln (§6) — an adjectival or pronominal attribute, the copula pw, an adverb or prepositional phrase, a particle, a suffix pronoun on the head, and a few others — with the qualification that almost none of them is strictly obligatory, since there is a way around each. Werning confirms the two clearest by counting: two bare nouns go direct nine times out of ten, a suffix pronoun on the head makes the construction indirect every time. Excluded here are 137 attestations: the 133 with an intervening determiner, none of which predates 1550 BCE, since an intervening determiner presupposes the definite article, and four with a suffix pronoun on the head. All 137 are indirect without exception.

Why a model and not a count

The obvious objection to any table of proportions in Egyptology is that the things being counted are not alike. Four problems in particular, all of which a hierarchical model is built to handle and a cross-table is not.

Everything is unbalanced. zꜣ “son” appears as head 1,005 times, ꞽwf “flesh” 642 times, nb “lord” 562; the median head is attested seven times, and the twenty commonest heads carry 42 per cent of all attestations. Some texts contribute two attestations, others two hundred. A raw proportion lets the big contributors decide, and “the corpus says” quietly means what the half-dozen largest sub-corpora say.

Everything is confounded with everything. A genre looks archaising until one notices it happens to use certain words; a lexeme looks conservative until one notices it occurs mainly in early texts. Counting one factor at a time cannot separate these. A model weighs them simultaneously and reports what is left of each when the others are held constant — which is why, for example, the animacy effect below is nearly invisible in the raw figures and clear once frequency is accounted for.

Rare things must not be over-read. The most useful property of a hierarchical model is one Egyptologists will recognise as an old methodological worry made precise. A lexeme attested twice, both times with the indirect genitive, does not get a rate of 100 per cent; it is pulled towards the corpus average, and how far depends on how much evidence it has. Lexemes with many attestations keep their own value, lexemes with few borrow strength from the rest. Nothing has to be thrown away for being rare, and nothing rare gets to shout.

Dates are spans, not years. The time axis in the model is a smooth curve fitted to the midpoints of dating ranges, which is a compromise, but a stated one.

The result is a hierarchical Bayesian logistic regression: separate levels for the head lemma, the dependent lemma and their combination, alongside text, sentence, sub-corpus and genre, plus a curve over time and a set of ordinary predictors. Fitted in Stan; the implementation is not the subject here.

Reading the numbers below

Two conventions, because the rest of the post uses them throughout.

Coefficients are in log-odds. They are additive on a scale that turns into percentages only once you know where you started. A coefficient of +0.33 means: if comparable cases sat at 40 per cent indirect, this feature moves them to about 48. The same +0.33 applied to cases already at 80 per cent would only move them to 85 — the scale compresses at the ends, which is exactly what you want when the thing being modelled is a proportion. Negative means “towards the direct genitive”, positive means “towards the indirect one”. As a rough guide: ±0.2 is small, ±1 is large, ±2 is overwhelming.

Intervals mean what they appear to mean. The brackets are 90 per cent credible intervals, and this is where the Bayesian framework earns its keep for philological work: “−0.99 [−1.65; −0.35]” says there is a 90 per cent probability that the value lies in that range, given model and data. That is the statement everyone wants to make and which a classical confidence interval does not license. When an interval covers zero, the data do not tell us the direction — which, as will become clear, happens more often here than one would like.

The head noun is what matters

Before the individual claims, one result that none of them states. Of everything the lexicon contributes to the choice, the combination of the two nouns accounts for roughly 57 per cent, the head for 36, and the dependent for only 7. The head is worth five times the dependent — Jansen-Winkeln’s thesis, confirmed. But the largest share belongs to neither noun separately. Which two words stand together matters more than either one does alone, and no property of either noun predicts it.

zꜣ “son” shows what that means in practice. Werning finds all nine kinship uses in the Book of Caverns direct and explains them as an old lexicalised compound, with Coptic ⲥⲓⲧⲉ “basilisk” from zꜣ tꜣ as the residue. Across the whole corpus it is a strong preference rather than a categorical one: 22 per cent of its 1,005 attestations are indirect, and they run through every period.

The number rule holds — for two formulas

Jansen-Winkeln’s clearest lexical observation concerns monosyllabic masculines, especially body parts: in the Pyramid Texts they take the indirect genitive in the singular and the direct one in the dual and plural. He does not confine the pattern to that corpus — he cites ḥm nï “the majesty of” and finds traces still in Late Middle Egyptian — and his list of lexemes ends with “and others”.

Half of this is right. Raw, the singular stands indirect in 40.6 per cent of cases, the plural in 40.6 — no difference at all — and the dual in 16.2. In the model, with everything else held constant, the dual sits at −0.99 [−1.65; −0.35] and the plural at +0.07 [−0.29; +0.41]. So the pull towards the direct genitive is substantial in the dual and absent in the plural: a refinement rather than a refutation, since paired body parts in the dual are the core of his examples.

But it is less even than that. Of 419 dual attestations, 252 belong to three head lemmas, and two of those are single fixed expressions: ꜥꜣ-wï= p:t “the two door leaves of the sky”, 99 attestations, and zḫn-wï= p:t, 89 attestations, which is one lemma pair. Both are direct almost without exception. Remove the three commonest dual heads and the dual stands at 35.9 per cent against the singular’s 40.6 — five points instead of twenty-four.

The model has a level for the lemma pair precisely so that formulas like these can be absorbed there rather than masquerading as grammar, and it half does: ꜥꜣ + p:t sits at −0.93, but with an interval of [−2.07; +0.20] that covers zero. With 99 attestations of one expression and almost no counter-evidence from the same pair in the singular, the model cannot fully decide whether what it sees is a property of the expression, of the lexeme, or of the dual. All three explanations fit.

The honest formulation is that the rule holds for certain lexicalised dual expressions rather than for the dual as a grammatical category. That is not a small distinction: it moves the phenomenon from morphology into the lexicon, which is where Jansen-Winkeln located the whole question in the first place.

Material and measure: strong, but not a rule

He also holds that expressions of material, length and content always take the indirect genitive; Werning finds the same for the genitivus materiae, which appears only indirectly on his semantic map.

The claim is not originally Jansen-Winkeln’s. He reports it from Erman and Osing, who separate two genitives — belonging and dependence on the one hand, qualifying specification (material, measure, content, kind) on the other — and he adds a morphological account: the second type goes back to a reversed, passive nisba nï, as in mdw nï z “the man’s word” against z nï mdw “a man of words” (§§9–11). On his view the second type should strictly speaking not be called a genitive at all.

The preference is strong and the exceptionlessness is not there. Material words as dependent stand indirect in 68.8 per cent of cases and measures in 72.2, against 40.0 elsewhere — but 29 material and 10 measure attestations stand in the direct genitive.

Whether the preference belongs to the class or to its eight members cannot be decided here: the class effect sits at +0.50 [−0.19; +1.19] and demonstrates nothing. What carries the class is three words: ṯḥn:t “faience, glass” at 94 per cent of 17 attestations, ḥmt “copper, ore” at 83 per cent of 24, and mḥ “cubit” at 73 per cent of 33. The second-largest material lexeme, bd “natron”, stands at 45 per cent of 22, barely above the corpus at large. “Material words” as a category carry nothing.

Werning’s animacy artefact does not appear here

Werning’s animacy finding is not really a claim about animacy. It is a claim about an artefact: the reversed relation he measures is, he argues, a by-product of cases like nb= ḏꜣḏꜣ “owner of the head”, where the possessum grammatically owns the possessor. Remove those and the relation nearly disappears. His raw figures put the reversal plainly: against Kammerzell’s expectation and against the iconicity principle, inalienable possession stands indirect in 35 per cent of his cases and alienable in only 15.

Checking this needs an animacy scale, which the TLA does not have — its dependent categories (divine name, toponym, title, royal name, personal name, common noun) are not one, because the common noun mixes people with things. So 200 dependent lexemes covering 88 per cent of attestations were coded by hand, on a sheet that deliberately did not show how those lexemes behave.

The test is a comparison of two runs, one with his class in the model and one without, since his claim is about distortion rather than about a value. The animacy coefficient moves by 0.002 — two thousandths of a log-odds unit, which is nothing. His class itself sits at −0.29 [−0.92; +0.35] and carries nothing either.

Animacy in this corpus points the ordinary way rather than the reversed one: +0.33 [+0.06; +0.60] for human and divine possessors among common nouns, which is about eight percentage points. And it is a good illustration of why one models rather than counts, because raw it is nearly invisible: 41.0 per cent indirect for animate possessors against 39.5 for inanimate. A point and a half.

What hides it is frequency. Animate common nouns are much more frequent words, and frequency pulls towards the direct genitive; the two effects nearly cancel. Compare like with like and the difference appears — in the middle frequency band, animate possessors stand indirect in 49.7 per cent of cases against 30.2 for inanimate ones. Twenty points. A cross-table would have reported “no animacy effect” and been wrong, and a cross-table split by frequency would have run out of data.

Share of indirect genitives for inanimate and for human or divine possessors, overall and in three frequency bands. Overall the two are within two points of each other; in the middle band the animate group is nearly twenty points higher. Frequency hides the animacy effect Share of indirect genitives among appellative (common-noun) dependents, split by how frequent the dependent lemma is. Raw proportions, no controls; bands are thirds of lemma frequency. inanimate possessor human or divine possessor 0% 20% 40% 60% 80% All appellative dependents 39.5% (n = 5,046) 41.0% (n = 4,416) Rare third 45.7% (n = 2,784) 48.4% (n = 380) Middle third 30.2% (n = 2,152) 49.7% (n = 1,579) Frequent third 68.2% (n = 110) 34.3% (n = 2,457) Taken together the two groups differ by 1.5 points; in the middle band, where both are well attested, by 19.5. Animate common nouns are much more frequent words, and frequency itself pulls towards the direct genitive; the two effects nearly cancel. Outlined bar: the frequent band holds only 110 inanimate attestations and carries nothing.
Figure 1: Appellative (common-noun) dependents only: 9,462 of the 12,274 appellative attestations, animals and undetermined lexemes left out. Bands are thirds of dependent-lemma frequency. Raw proportions. The top band holds only 110 inanimate attestations, which is why it points the other way.

One caveat on the negative result. Because every head lemma has its own effect in the model, nb, ḥqꜣ and nzw are already accounted for individually before Werning’s class is switched on. What is tested here is whether the class distorts anything beyond that. It does not — but that is not the same as showing his correction is unnecessary in a simple cross-table, where it may well be needed.

Archaising belongs to compositions, not to genres

Werning reads the 80 per cent direct genitives of the Book of Caverns as deliberate archaising. The contrast that makes the reading plausible is with his contemporaries: across the New Kingdom the corpus sits at 51.7 per cent indirect, so a text at 20 per cent is far outside the range of what was usual when it was composed.

What the Book of Caverns gets wrong is not only the proportion. Jansen-Winkeln stresses that in Old and Middle Egyptian the direct genitive is not a compound but a free word group: the dependent can be expanded by a suffix, a further genitive, an adjective or a relative phrase, and the sequence can be broken (§4). Werning finds in the Book of Caverns 14 clear and 23 disputed cases of directly nested genitives, a pattern all but unattested in genuine Old and Middle Egyptian, and concludes that its authors no longer understood the direct genitive as a phonologically marked compound but as bare juxtaposition. The rule he reconstructs for their mental translation — delete the article, inflect the connector or delete it — contains no phonology at all.

That is a prediction one can check: religious compositions of the New Kingdom should stand markedly more direct than contemporaries once date, lexicon and grammar are held constant.

They do, and strongly. Religious and funerary texts of the New Kingdom sit at −1.73 [−2.14; −1.33] against everything else of their time, with the Book of the Dead at −1.31.

Share of indirect genitives by New Kingdom sub-corpus, from medical texts at 69.4 per cent down to letters at 20.8 per cent. Indirect genitive by sub-corpus, New Kingdom Share of attestations using the indirect construction, 1550–1070 BCE. Sub-corpora with at least 150 attestations (4,619 of 4,856). Dashed line: New Kingdom average, 51.7%. 0% 20% 40% 60% Medical texts 69.4% n = 725 Netherworld books 59.7% n = 278 Amarna texts 58.4% n = 161 Magical texts 58.4% n = 425 Ramesside inscriptions 53.1% n = 569 Literary texts 52.0% n = 763 Magico-medical (heka) 50.7% n = 290 Book of the Dead 42.3% n = 665 Historical-biographical 40.1% n = 551 Letters 20.8% n = 192
Figure 2: Sub-corpora of the New Kingdom with at least 150 attestations, 4,619 of 4,856. Proportions are raw; the ordering survives holding date, lexicon and grammar constant.

The spread is nearly fifty points, and none of it is chronological — this is all the same five centuries. Technical prose sits at the indirect end: medical texts 69.4 per cent, magical texts 58.4. Everyday documents sit at the direct end, letters lowest of all at 20.8. The Book of the Dead, at 42.3 per cent, is nearly ten points below literary texts of the same period. That is register: the same language, and writers choosing differently according to what they were writing.

The prediction fails for Werning’s own material, though, and for a reason worth knowing. The netherworld books of the New Kingdom sit at +0.10 [−0.37; +0.59] — on the average, nothing archaising about them. And the Book of Caverns itself is in the TLA with six text pieces, five dated to the 30th Dynasty and one to Psammetichus I, because the TLA dates by witness and the surviving copies come from TT 33 and later. Only two of its attestations reach this study at all. His text is effectively not in this corpus.

What that leaves is sharper than the original question. Archaising is not a property of religious texts in general but of particular compositions: the Book of the Dead holds on to the direct genitive, the netherworld books do not — same period, same lexicon, same wider category of text.

The largest effect has no explanation

One thing the model turns up is larger than anything discussed so far, and the corpus cannot settle what lies behind it.

Because the number rule is lexical rather than general, each head lemma gets its own slope for “not singular” — its own answer to the question of what the plural does to it. The spread of those slopes is 2.11 in log-odds, the largest single quantity anywhere in the analysis: larger than any group level, larger than date, larger than any fixed effect. And it runs in both directions. ꞽb= nṯr “the god’s heart” is indirect in 70 per cent of its attestations; ꞽb-w= nṯr-w “the gods’ hearts” in none of its fifteen plural attestations. ꜥ “arm” is indirect in 72 per cent of singulars and 2 per cent of duals. ḥꜣb “festival” goes the opposite way: 22 per cent in the singular, 83 in the plural.

For 22 head lemmas, the share of indirect genitives in the singular on the left and in the dual or plural on the right. Body and soul parts fall steeply, other lemmas rise. What the plural does to a head noun depends on the noun Share of indirect genitives in the singular and in the dual or plural, for the 22 head lemmas with at least ten attestations in both. Raw proportions, no controls. 0% 25% 50% 75% 100% singular dual / plural ꞽb 76% bꜣ 74% ꜥ 72% ḥꜥ-w 68% sḫr 65% ꜥꞽ 33% ꜥrf 27% ḥꜣb 22% ḫꜣs:t 12% ḥꜣb 83% ꜥꞽ 54% ḫꜣs:t 47% ꜥrf 41% ḥꜥ-w 33% sḫr 32% bꜣ 17% ꞽb 9% ꜥ 2% Blue: body and soul parts, pulled towards the direct genitive in the plural.Orange: lemmas pulled the other way. Grey: the remaining thirteen.
Figure 3: Head lemmas with at least ten attestations in both numbers. Raw proportions, no controls; blue marks body and soul parts, orange the lemmas that move the other way.

The first thing to rule out is that this is an artefact of the shrinkage described earlier — that the model is inventing variation among lexemes it knows nothing about. It is not: 412 of 625 head lemmas have no non-singular attestation at all, and their slopes have a spread of 0.02 against 1.47 for the rest. The model pulls the uninformed ones to zero exactly as it should, and the 2.11 comes from the roughly 35 lemmas that have real evidence.

The cleanest demonstration that the effect is linguistic dispenses with the model entirely. Of the 42 lemma pairs attested at least five times in both numbers, the average difference in proportion between singular and non-singular is 24 percentage points, and eleven differ by more than 30. Same head, same dependent, different number, different construction.

The obvious explanation is alienability. The lemmas pulling towards the direct genitive in the plural are, on the whole, body parts and soul parts — ꜥ, ꞽb, qs “bone”, bꜣ, dp — and those pulling the other way are alienable things and events. It would be a striking result, because Werning finds no alienability split as a main effect: it would exist only in interaction with number.

It does not survive testing. Coding 213 head lemmas by hand and adding alienability as a predictor on the slope gives −0.75 [−1.33; −0.15] across all coded lemmas. But thirty-five of those lemmas are the ones that suggested the hypothesis in the first place. Restricted to the other 178, it is −0.47 [−1.14; +0.20]: the right direction, about 87 per cent of the probability mass below zero, and an interval that includes nothing happening. And even in the generous version it accounts for about 8 per cent of the variation it is meant to explain; in the honest version, 2.

A sharper test would need more duals and plurals than exist. There are 1,622 in 16,812 attestations, 9.6 per cent, spread over 213 lemmas.

There is one candidate explanation older than alienability, and it comes from Jansen-Winkeln himself. He takes the complementarity of the monosyllables — indirect in the singular, direct in the dual and plural — to have morphological and phonological rather than semantic causes, and in a footnote he offers a mechanism: a monosyllable necessarily has a short stressed vowel in the singular, so there is no long-to-short reduction for the status constructus to make, whereas the plural and the feminine have one, and in the dual a reduction would at least be easier (§8 with n. 78). He marks this as a conjecture. The lemmas with the steepest negative slopes here are largely his monosyllables. What stands in the way of testing it is that every head lemma would have to be classified by syllable structure and by its behaviour in the status constructus, and the TLA annotation supplies neither. The explanation that fits best is the one this corpus cannot be asked about.

So the largest thing this analysis measures is real, is lexical, and is not explained here.

A sentence from 2000

Having described how lexicon and morphology steer the choice, Jansen-Winkeln writes (§8):

Diese Regeln im einzelnen (und für die verschiedenen Epochen) zu entschlüsseln, würde > umfangreiche Statistiken erfordern. > > To decode these rules in detail, and for the various periods, would require extensive > statistics.

The sentence is twenty-six years old. He asks for extensive statistics. What they need has come into being since: a digital, lemmatised corpus, and the computing power for a Bayesian model that gives each lemma an effect of its own. This post works through one, for the part of the material where both constructions were genuinely available. The commonest result is that the question cannot be decided here. That, too, is an answer.


16,812 attestations from TLA v20, all periods. Hierarchical Bayesian logistic regression in Stan: crossed levels for head lemma, dependent lemma and lemma pair, alongside text, sentence, sub-corpus and genre, a cubic spline over date, and predictors on the lexeme levels; non-centred parameterisation throughout, four chains, no divergences. Code, coding sheets and full coefficient tables are not published; I am glad to share them on request. This is a side project and the findings are provisional.