Template:etymon
- The following documentation is located at Template:etymon/documentation. [edit]
- Useful links: subpage list • links • redirects • transclusions • errors (parser/module) • sandbox
This template is used to indicate the entry term's immediate ancestor (etymon).
Note that generally, an "etymon" can refer to any kind of etymological ancestor, but on this page, "etymon" will be used specifically for an immediate ancestor. For example, the etymon of English nexus is Latin nexus, and the etymon of Latin nexus is Latin nectō. Terms can have multiple etymons, so the etymons of English toothbrush are both tooth and brush.
Current features: automatic categorization, tree generation, text generation, reference support.
Usage
Suppose that the etymon of English sow (“female pig”) is set to Middle English sowe, and the etymon of sowe is set to Old English sugu. The template is then able to intelligently connect sow and sugu, even though they are two steps apart. The template thus reduces duplication across entries by making it unnecessary to manually specify that sow and sugu are connected.
Etymon IDs
Etymon IDs are used to distinguish between multiple etymologies for the same term (homographs), in order to ensure that the template is able to traverse these etymological chains without getting confused.
This mirrors the practice of the Oxford English Dictionary. For example, the two main senses of sow are identified as 9086735656 (“to plant seeds”) and 9710708293 (“female pig”). On Wiktionary, the identifier should be a word or short phrase which summarizes the definition of the term rather than a meaningless string of numbers. In the case of sow, the two IDs might be plant seeds and female pig.
For terms with only one etymology, no ID is required either in the template's |id= parameter or in the etymon parameters. If a page has multiple etymon templates for the same language, all of them must have unique IDs—the template will throw an error if it finds multiple etymon templates where at least one is missing an ID.
Parameters
The template takes the following parameters:
|1=(required)- The language code of the current entry.
|id=- The etymology ID. This parameter creates an anchor to the current section. For example, if the template at English father is
{{etymon|en|id=male parent}}, this etymology section is directly linked to by father#English: male parent. The ID must have at least two characters and must not be the same as the page title. This parameter is optional when there is only one etymology for the term, but required when there are multiple etymologies. |title=- This parameter manually overrides the current page title. For example, if an etymology tree is created at Latin pōnō (located at pono) it is necessary to specify
|title=pōnōto ensure that the macrons are displayed. |pos=- Indicates the part of speech of the current entry, which allows descendants to intelligently categorize themselves. Usually only necessary for Proto-Indo-European. Current allowed values are:
prefix,suffix,interfix,infix,root,word. |exnihilo=- If set to anything, adds the entry to the category
<language> terms coined ex nihilo. |etydate=- Specifies the date or time period when the term was first attested. Supports inline modifiers for references and formatting. For example:
etydate=1500s<ref:or{{R:OED}}>etydate=late 14c<nocap>. Ranges can also be given, e.g.etydate=r,1501,1558<ref:{{R:pl:SXVI|artykuł|5767}}> |rfe=- Adds a request for etymology notice. Can be set to
1for a standard notice, or to custom text for a specific request. Supports inline modifiers:<nocat>,<noes>,<box>,<sort:key>,<y:year>,<m:month>,<fragment:section>,<section:name>. For example:|rfe=1,rfe=Need more details<nocat>, orrfe=1<box><sort:custom>. |nl=- If set to anything, consider the entry a non-lemma.
|nodot=- If set to anything, removes the final period from the generated etymology text. Useful when combining with other text or templates.
|nocat=- If set to anything, suppresses automatic categorization.
|tree=- If set to anything, displays an etymology tree.
|text=- If set to anything, displays some text describing the etymology. The text modes are:
+— single step (immediate parent only)++— all steps (full etymology chain)*— to the nearest blue link (stops when it reaches an existing entry)- A number (e.g.,
3) — show up to that many steps :langcode(e.g.,:ar) — show steps until reaching a term in the specified language:langcode*(e.g.,:ota*) — stop at the specified language if it's a blue link, otherwise continue to the first blue link
- For more information, see § Text.
|2=,|3=, ...- These are the etymon parameters. Each etymon parameter can be either an etymon or a derivation keyword. There can be any number of etymon parameters.
Etymons
Etymons must be written using the following format: languagecode:term<id:identifier><ref:reference>. For example: en:pay<id:give money> represents English pay (etymology 1). As a shortcut, it is possible to omit the language code (i.e., just writing pay<id:give money>). In this case, the template will assume that the language is the same as the one set in |1=.
The ID is not mandatory - simple etymons can be written as just languagecode:term (e.g., en:father) or term (e.g., father when the language matches parameter 1). However, if there are multiple etymologies for the same term on the target page, an ID must be specified to distinguish between them.
To suppress the term display (show only the language), use - as the term, e.g., la:-.
Additional inline modifiers can be used:
<id:identifier>- specifies the etymology ID. Prefix with!(e.g.,<id:!forced>) to force the ID even when disambiguation would normally not require it.<t:gloss>- provides gloss/translation<tr:transliteration>- provides transliteration<ts:transcription>- provides transcription<alt:alternative display>- alternative display text. Use this to suppress linking when referring to proto-language entries with a single descendant, which, according to policy, should not be created.<pos:part of speech>- part of speech annotation<ety:inline etymology>- inline etymology for redlinks (see section below)<unc>- when set, marks the corresponding step or element as uncertain. Note that discredited, dubious, or speculative etymologies should not be added to the template at all.<ref:reference text>- adds references (see References section)<aftype:type>- specifies the affix type when used with:af. Valid values:prefix/pre,suffix/suf,infix/in,interfix/inter,circumfix/circum,non-affix/naf,root.<postype:type>- specifies whether a term should be treated as arootorwordfor categorization purposes. Overrides the default detected from the linked page's|pos=.<bor>- when used on a term under:af, adds borrowing categories for that specific component (for affixed terms where one component is borrowed).<lbor>and<slbor>are also valid term modifiers.<senseid>- used for senseid's, see below on semantic loans. Multiple can be given by separating the id's with !!, e.g. <senseid:sense1!!sense2>.
Inline etymology
For terms that don't have their own entry (redlinks) or are missing an etymon template, you can specify the etymology inline using the <ety:...> modifier. This allows the tree to continue beyond redlinks.
The syntax is: <ety:keyword<etymon1><etymon2>...>
For example:
enm:example<ety:inh<ang:example>>- specifies that the Middle English term is inherited from Old Englishla:verbum<ety:der<grc:ῥῆμα>>- specifies that the Latin term is derived from Greek
The inline etymology is only used when the target page is a redlink or is missing an etymon template. If the target page has a valid etymon template, the inline etymology is ignored (and a tracking category is added).
Derivation keywords
Each derivation keyword applies to all the etymons that follow it. Derivation keywords must be prefixed with a colon (:). A derivation keyword is reset by another derivation keyword. For example, |:inh|etymon1|etymon2|:bor|etymon3|etymon4 means that a term is inherited from both etymon1 and etymon2 and also borrowed from both etymon3 and etymon4.
Using an unknown keyword (i.e. one not listed below) produces an error message.
Keyword modifiers
Keywords can have inline modifiers attached using angle brackets:
<unc>: marks the derivation as uncertain; for example,:bor<unc>means "possibly borrowed from".<ref:reference text>: adds a reference to the keyword (same syntax as etymon references).<conj:conjunction>: specifies a custom conjunction when there are multiple etymons under the same keyword (default is "or"); Examples::bor<conj:and/or>|es:cruzado|pt:cruzadoproduces "Borrowed from Spanish cruzado and/or Portuguese cruzado.":bor<conj:and>|etymon1|etymon2produces "...etymon1 and etymon2."
<surf>- used for surface analysis.
Basic derivation
:from: (default) unspecified derivation type within a language; corresponds to{{from}}; only valid for same-language derivations.:derived(or:derin short): used when a term is derived from another language; corresponds to{{derived}}.:inherited(or:inh): used when a term comes directly from the parent language unchanged; corresponds to{{inherited}}. The template validates that the source language is a valid ancestor.:uder(short for undefined derivation; no long form yet): used when the exact relationship is unclear; adds to "undefined derivations" category.
Borrowings
:bor(alias:borrowed): used when a term is borrowed from another language directly; corresponds to{{borrowed}}.:lbor(short for learned borrowing): used for a learned borrowing, that is a term is borrowed intentionally rather than through normal language contact; corresponds to{{learned borrowing}}.:slbor(short for semi-learned borrowing): used for a learned borrowing which is reshaped somewhat; corresponds to{{semi-learned borrowing}}.:obor(short for orthographic borrowing): used for orthographic borrowings; corresponds to{{orthographic borrowing}}.:ubor(short for unadapted borrowing): used for direct borrowings that retain original orthography; corresponds to{{unadapted borrowing}}.
Semantic relationships
:calque(aliases:calor:clq): used for calques; corresponds to{{calque}}; non-transitive (does not follow the chain further).:partial calque(or:pcal): used for partial calques; corresponds to{{partial calque}}; non-transitive.:semantic loan(or:sl): used for semantic loans; corresponds to{{semantic loan}}; non-transitive; the loaned sense may be linked using <senseid>, so long as{{senseid}}is present near the appropriate definition.:influence(no short form yet): used when a term is influenced in some way by another (e.g. the modern meaning of English discomfit is influenced by the unrelated word discomfort); non-transitive.
Word formation
:affix(or:af): used for compounds, affixation, and any other template whose output contains a+between two terms; parameter order matters:|:af|etymon1|etymon2means etymon1 + etymon2; corresponds to{{affix}},{{compound}}, and others; generates affix categories automatically. Keywords like ``:af`` or ``:deverbal`` can be marked with <surf> for surface analysis.:afeq(short for "equivalent affix"): used as a "hidden{{af}}", meaning that etymons with this keyword shall be invisible in tree and text output, but their respective categories shall still be generated; as an example, childhood, from Middle English childhode (pronounced cheeld-hohd), is equivalent to child + -hood.:blend: used for blends; corresponds to{{blend}}.:reduplication(or:redup): used for reduplications; corresponds to{{reduplication}}.:abbreviation(or:abbr): used for abbreviations; corresponds to{{abbreviation}}.:syllabic abbreviation(or:sylabbr): used for syllabic abbreviations; corresponds to{{syllabic abbreviation}}.:acronym(alias:acro): used for acronyms; corresponds to{{acronym}}.:initialism(alias:init): used for initialisms; corresponds to{{initialism}}.:clipping(or:clip): used for clippings; corresponds to{{clipping}}.:ellipsis(alias:ellip): used for elliptical forms; corresponds to{{ellipsis}}.:univerbation(or:univ): used for univerbation; corresponds to{{univerbation}}.:back-formation(alias:bf): used for back-formations; corresponds to{{back-formation}}.:apheretic(alias:apheresisor:aphetic): used for apheretic forms.:denominal(alias:denom): used for denominals; corresponds to{{denominal verb}}.:deverbal: used for deverbals; corresponds to{{deverbal}}.:sa-af: used for Sanskritic formations.:vrd-af: used for Sanskrit vṛddhi derivatives with an affix. See:vrdbelow.
Other
:transliteration(or:translit): used when a term is transliterated from another script; corresponds to{{transliteration}}.:vrd(short for "vṛddhi derivative"): used for Sanskrit vṛddhi derivatives.:root: marks the following etymons as roots for categorization purposes. The etymons will be categorized under "terms belonging to the root X" or "terms derived from the root X". This keyword is invisible in tree and text output.
References
References can be added to etymons using inline modifiers with the syntax <ref:reference text>. Multiple references can be separated by !!! (with spaces around the exclamation marks). References support the same syntax as other Wiktionary reference templates, including named references and groups (see Module:references). References are only displayed when |text= is set and only on the same entry they are defined, not on any descendants.
Examples:
la:verbum<ref:- single reference{{R:L&S}}>- Equivalent to
<ref>{{R:L&S}}</ref>
- Equivalent to
la:verbum<ref:- multiple references with naming{{R:L&S}}<<name:LS>> !!!{{R:Gaffiot}}>- Equivalent to
<ref name="LS">{{R:L&S}}</ref><ref>{{R:Gaffiot}}</ref>
- Equivalent to
la:verbum<ref:<<name:LS>>>- reference to previously named reference- Equivalent to
<ref name="LS"/>
- Equivalent to
la:verbum<ref:- named reference with group{{R:L&S}}<<name:LS>><<group:etymology>>>- Equivalent to
<ref name="LS" group="etymology">{{R:L&S}}</ref>
- Equivalent to
Trees
If the parameter |tree=1 is set, an etymology tree is inserted.
Per an April 2024 vote, each language community decides when it is appropriate to show a tree on a particular entry.[1] Additionally, trees should not be displayed for clear open compounds like United States of America.[2]
Keywords with no_child_categories (such as :calque, :sl, :pcal, :influence) will have their subtrees hidden in the tree display, as these represent non-transitive relationships where the further etymology is not directly relevant. Similarly, duplicate nodes (when the same term appears multiple times in a tree) do not re-display their ancestry. In both cases, an L-shaped dotted connector indicates hidden ancestry, and duplicate nodes are additionally styled with dashed borders.
Text
The |text= parameter may be set to automatically generate a text description from the given data. The parameter should only be used in entries under the languages whose editing communities have explicitly approved it and as long as this does not imply the removal of valuable information which cannot be produced by the module.[3]
The following languages and families have allowed its usage. Any language in an allowed family (e.g. Indo-Aryan for family code inc) is also allowed.
| Code / family | Languages |
|---|---|
| Language codes | amf (Hamer-Banna), bg (Bulgarian), bnt-sab-pro (Proto-Sabaki), btk-pro (Proto-Batak), cmc-pro (Proto-Chamic), cs (Czech), dru-pro (Proto-Rukai), en (English), eo (Esperanto), es (Spanish), ext (Extremaduran), fa (Persian), hsb (Upper Sorbian), iir-pro (Proto-Indo-Iranian), jbo (Lojban), jdt (Judeo-Tat), la (Latin), map-ata-pro (Proto-Atayalic), map-pro (Proto-Austronesian), mul (Translingual), ota (Ottoman Turkish), phi-kal-pro (Proto-Kalamian), phi-pro (Proto-Philippine), poz-btk-pro (Proto-Bungku-Tolaki), poz-cet-pro (Proto-Central-Eastern Malayo-Polynesian), poz-hce-pro (Proto-Halmahera-Cenderawasih), poz-lgx-pro (Proto-Lampungic), poz-mcm-pro (Proto-Malayo-Chamic), poz-mic-pro (Proto-Micronesian), poz-mly-pro (Proto-Malayic), poz-msa-pro (Proto-Malayo-Sumbawan), poz-oce-pro (Proto-Oceanic), poz-pep-pro (Proto-Eastern Polynesian), poz-pnp-pro (Proto-Nuclear Polynesian), poz-pol-pro (Proto-Polynesian), poz-pro (Proto-Malayo-Polynesian), poz-ssw-pro (Proto-South Sulawesi), poz-swa-pro (Proto-North Sarawak), pqe-pro (Proto-Eastern Malayo-Polynesian), ps (Pashto), ro (Romanian), sk (Slovak), sla-pro (Proto-Slavic), sw (Swahili), tg (Tajik), tl (Tagalog), tr (Turkish), uk (Ukrainian), uz (Uzbek), zle-ono (Old Novgorodian), zlw-ocs (Old Czech), zlw-osk (Old Slovak)
|
ber (Berber) |
auj (Awjila), ber-fog (Fogaha), ber-pro (Proto-Berber), ber-zuw (Zuwara), cnu (Chenoua), gha (Ghadames), gho (Ghomara), gnc (Guanche), jbn (Nefusa), kab (Kabyle), mzb (Northern Saharan Berber), rif (Tarifit), sds (Tunisian Berber), shi (Tashelhit), shy (Tachawit), siz (Siwi), sjs (Senhaja de Srair), swn (Sokna), tez (Tetserret), tmh (Tuareg), tzm (Central Atlas Tamazight), zen (Zenaga), zgh (Moroccan Amazigh)
|
dra (Dravidian) |
aaf (Aranadan), all (Allar), bfq (Badaga), brh (Brahui), brw (Bellari), cde (Chenchu), ctt (Wayanad Chetti), daq (Dandami Maria), dra-bry (Beary), dra-cen, dra-cen-pro (Proto-Central Dravidian), dra-gki, dra-gon, dra-imd, dra-kan, dra-kki, dra-kml, dra-knk, dra-kod, dra-kor, dra-mal, dra-mdy, dra-mkn (Middle Kannada), dra-mlo, dra-mur, dra-nor, dra-nor-pro (Proto-North Dravidian), dra-okn (Old Kannada), dra-ote (Old Telugu), dra-pgd, dra-pro (Proto-Dravidian), dra-sdo, dra-sdo-pro (Proto-South Dravidian I), dra-sdt, dra-sdt-pro (Proto-South Dravidian II), dra-sou, dra-sou-pro (Proto-South Dravidian), dra-tam, dra-tel, dra-tkd, dra-tkn, dra-tkt, dra-tlk, dra-tml, emu (Eastern Muria), era (Eravallan), fmu (Far Western Muria), gau (Kondekor), gdb (Ollari), gon (Gondi), hca (Andaman Creole Hindi), hoy (Holiya), ima (Mala Malasar), iru (Irula), kej (Kadar), kep (Kaikadi), kev (Kanikkaran), kfa (Kodava), kfb (Kolami), kfc (Konda-Dora), kfd (Korra Koraga), kfe (Kota (India)), kff (Koya), kfg (Kudiya), kfh (Kurichiya), kfi (Kannada Kurumba), kmj (Kumarbhag Paharia), kn (Kannada), kpb (Mullu Kurumba), kru (Kurux), kwx (Khirwar), kxu (Kui (India)), kxv (Kuvi), mha (Manda (India)), mjo (Malankuravan), mjp (Malapandaram), mjq (Malaryan), mjr (Malavedan), mjt (Sawriya Paharia), mju (Manna-Dora), mjv (Mannan), ml (Malayalam), mrr (Hill Maria), mut (Western Muria), muv (Muthuvan), nbg (Nagarchal), nit (Southeastern Kolami), oty (Old Tamil), pcf (Paliyan), pcg (Paniya), pch (Pardhan), pci (Duruwa), peg (Pengo), pkr (Attapady Kurumba), ptq (Pattapu), pty (Pathiya), sle (Sholaga), ta (Tamil), tcx (Toda), tcy (Tulu), te (Telugu), thn (Thachanadan), udg (Muduga), ull (Ullatan), url (Urali), vis (Vishavan), vmd (Mudu Koraga), wbq (Waddar), wkb (Kumbaran), wkl (Kalanadi), wku (Kunduvadi), xis (Kisan), xua (Alu Kurumba), xub (Betta Kurumba), xuj (Jennu Kurumba), yea (Ravula), yeu (Yerukula), ymr (Malasar)
|
iir-nur (Nuristani) |
ask (Ashkun), bsh (Kamkata-viri), iir-nur-pro (Proto-Nuristani), nur-nor, nur-sou, prn (Prasuni), trm (Tregami), wbk (Waigali)
|
inc (Indo-Aryan) |
aee (Northeast Pashayi), aeq (Aer), ahr (Ahirani), anp (Angika), anr (Andh), as (Assamese), awa (Awadhi), bdv (Bodo Parja), bfb (Pauri Bareli), bfr (Bazigar), bfy (Bagheli), bfz (Mahasu Pahari), bgc (Haryanvi), bgd (Rathwi Bareli), bge (Bauria), bgq (Bagri), bgw (Bhatri), bh (Bihari), bha (Bharia), bhb (Bhili), bhd (Bhadrawahi), bhe (Bhaya), bhi (Bhilali), bho (Bhojpuri), bht (Bhattiyali), bhu (Bhunjia), bhx (Bhalay), bjj (Kannauji), bkk (Brokskat), bmj (Bote-Majhi), bn (Bengali), bns (Bundeli), bpx (Palya Bareli), bpy (Bishnupriya Manipuri), bra (Braj), btv (Bateri), ccp (Chakma), cdh (Chambeali), cdi (Chodri), cdj (Churahi), cih (Chinali), clh (Chilisso), ctg (Chittagonian), dcc (Deccani), dgo (Hindi Dogri), dhd (Dhundhari), dhn (Dhanki), dho (Dhodia), dhw (Danuwar), dmk (Domaaki), dml (Dameli), doi (Dogri), dry (Darai), dso (Desiya), dty (Doteli), dub (Dubli), duh (Dungra Bhil), dv (Dhivehi), dwz (Dewas Rai), emx (Erromintxela), gas (Adiwasi Garasia), gbk (Gaddi), gbl (Gamit), gbm (Garhwali), gdx (Godwari), ggg (Gurgula), ghr (Ghera), gig (Goaria), gjk (Kachi Koli), gju (Gojri), glh (Northwest Pashayi), goj (Gowlan), gra (Rajput Garasia), gu (Gujarati), gwc (Kalami), gwf (Gowro), gwt (Gawar-Bati), haj (Hajong), hca (Andaman Creole Hindi), hi (Hindi), hif (Fiji Hindi), hii (Hinduri), him, hkh (Pogali), hlb (Halbi), hnd (Southern Hindko), hne (Chhattisgarhi), hno (Northern Hindko), hns (Caribbean Hindustani), hoj (Hadoti), inc-apa (Apabhramsa), inc-ash (Ashokan Prakrit), inc-bas, inc-bhi, inc-bih, inc-cen, inc-chi, inc-dar, inc-dng, inc-dng-pro (Proto-Dangari), inc-dre, inc-eas, inc-hal, inc-hie, inc-hiw, inc-hnd, inc-ins, inc-kam (Kamarupi Prakrit), inc-kas, inc-kho (Kholosi), inc-koh, inc-krd, inc-krd-pro (Proto-Kamta), inc-kun, inc-mas (Middle Assamese), inc-mbn (Middle Bengali), inc-mgu (Middle Gujarati), inc-mid, inc-mor (Middle Odia), inc-nor, inc-nwe, inc-oas (Early Assamese), inc-oaw (Old Awadhi), inc-obn (Old Bengali), inc-ogu (Old Gujarati), inc-ohi (Old Hindi), inc-old, inc-oor (Old Odia), inc-opa (Old Punjabi), inc-pah, inc-pan, inc-pas, inc-pro (Proto-Indo-Aryan), inc-rom, inc-shn, inc-snd, inc-sou, inc-tha, inc-wes, jat (Jakati), jdg (Jadgali), jml (Jumli), jnd (Jandavra), jns (Jaunsari), kbu (Kabutra), keq (Kamar), kex (Kukna), key (Kupia), kfr (Kachchi), kfs (Bilaspuri), kft (Kanjari), kfu (Katkari), kfv (Kurmukar), kfx (Kullu Pahari), kfy (Kumaoni), khn (Khandeshi), khw (Khowar), kjo (Harijan Kinnauri), kls (Kalasha), kok (Konkani), kra (Kumhali), ks (Kashmiri), ksy (Kharia Thar), kvx (Parkari Koli), kxp (Wadiyara Koli), kyv (Kayort), kyw (Kudmali), lah (Lahnda), lmn (Lambadi), lss (Lasi), luv (Luwati), mag (Magahi), mai (Maithili), mby (Memoni), mjl (Mandeali), mjz (Majhi), mkb (Mal Paharia), mki (Dhatki), mr (Marathi), mtr (Mewari), mup (Malvi), mvy (Indus Kohistani), mwr (Marwari), nag (Naga Pidgin), ne (Nepali), nhh (Nahari), nli (Grangali), nlm (Mankiyali), nlx (Nahali), noe (Nimadi), noi (Noiri), oak (Noakhali), odk (Od), omr (Old Marathi), or (Odia), ort (Adivasi Odia), pa (Punjabi), paq (Parya), pcl (Pardhi), pgd (Gandhari), pgg (Pangwali), phd (Phudagi), phl (Palula), phr (Pahari-Potwari), pi (Pali), plk (Kohistani Shina), pra (Prakrit), pra-niy (Niya Prakrit), psh (Southwest Pashayi), psi (Southeast Pashayi), pwr (Powari), raj, rge (Romano-Greek), rhg (Rohingya), rjs (Rajbanshi), rkt (Kamta), rmc (Carpathian Romani), rmd (Traveller Danish), rme (Angloromani), rmf (Kalo Finnish Romani), rmg (Traveller Norwegian), rmi (Lomavren), rml (Baltic Romani), rmn (Balkan Romani), rmo (Sinte Romani), rmq (Caló), rmt (Domari), rmu (Tavringer Romani), rmw (Welsh Romani), rmy (Vlax Romani), rom (Romani), rsb (Romano-Serbian), rtw (Rathawi), sa (Sanskrit), saz (Saurashtra), sbn (Sindhi Bhil), sck (Sadri), scl (Shina), sd (Sindhi), sdg (Savi), sdr (Oraon Sadri), shd (Kundal Shahi), si (Sinhalese), sjp (Surjapuri), skr (Saraiki), smm (Musasa), smv (Samvedi), soi (Sonha), spv (Sambalpuri), srx (Sirmauri), ssi (Sansi), sts (Shumashti), syl (Sylheti), tdb (Panchpargania), the (Chitwania Tharu), thl (Dangaura Tharu), thq (Kochila Tharu), thr (Rana Tharu), tkb (Buksa), tkt (Kathoriya Tharu), tnv (Tanchangya), tra (Tirahi), trl (Traveller Scottish), trw (Torwali), ur (Urdu), ush (Ushojo), vaa (Vaagri Booli), vah (Varhadi), vav (Varli), vgr (Vaghri), vjk (Bajjika), wbr (Wagdi), wsv (Wotapuri-Katarqalai), wtm (Mewati), xhe (Khetrani), xka (Kalkoti), xnr (Kangri), zrg (Mirgan)
|
mun (Munda) |
agi (Agariya), asr (Asuri), bfw (Bondo), bix (Bijori), biy (Birhor), cdz (Koda), ekl (Kolhe), gaq (Gata'), gbj (Bodo Gadaba), hoc (Ho), jun (Juang), juy (Juray), kfp (Korwa), kfq (Korku), khr (Kharia), ksz (Kodaku), lbm (Lodhi), mjx (Mahali), mun-pro (Proto-Munda), pcj (Parenga), sat (Santali), srb (Sora), trd (Turi), unr (Mundari), unx (Munda)
|
roa-gap (Galician-Portuguese) |
aoa (Angolar), ccd (Cafundó), cri (Sãotomense), crp-mpp (Macau Pidgin Portuguese), drc (Minderico), fab (Annobonese), fax (Fala), gl (Galician), idb (Indo-Portuguese), kea (Kabuverdianu), mcm (Kristang), mzs (Macanese), pap (Papiamentu), pov (Guinea-Bissau Creole), pre (Principense), pt (Portuguese), rmq (Caló), roa-opt (Old Galician-Portuguese), srm (Saramaccan), tvy (Timor Pidgin), vkp (Korlai Creole Portuguese)
|
sem-ara (Aramaic) |
aii (Assyrian Neo-Aramaic), aij (Lishanid Noshan), amw (Western Neo-Aramaic), arc (Aramaic), bhn (Bohtan Neo-Aramaic), bjf (Barzani Jewish Neo-Aramaic), cld (Chaldean Neo-Aramaic), hrt (Hértevin), huy (Hulaulá), kqd (Koy Sanjaq Surat), lhs (Mlahsö), lsd (Lishana Deni), mid (Mandaic), myz (Classical Mandaic), sam (Samaritan Aramaic), sem-are, sem-arw, sem-ase, sem-cna, sem-nna, syc (Classical Syriac), syn (Senaya), trg (Lishán Didán), tru (Turoyo), xrm (Armazic)
|
sem-arb (Arabic) |
abh (Tajiki Arabic), abv (Baharna Arabic), acm (Iraqi Arabic), acw (Hijazi Arabic), acx (Omani Arabic), acy (Cypriot Arabic), adf (Dhofari Arabic), aeb (Tunisian Arabic), afb (Gulf Arabic), ajp (South Levantine Arabic), apc (North Levantine Arabic), apd (Sudanese Arabic), ar (Arabic), arq (Algerian Arabic), ars (Najdi Arabic), ary (Moroccan Arabic), arz (Egyptian Arabic), auz (Uzbeki Arabic), ayl (Libyan Arabic), ayn (Yemeni Arabic), ayp (North Mesopotamian Arabic), kcn (Nubi), mey (Hassaniya Arabic), mt (Maltese), pga (Juba Arabic), shu (Chadian Arabic), sqr (Siculo-Arabic), ssh (Shihhi Arabic), xaa (Andalusian Arabic)
|
tup (Tupian) |
aan (Anambé), adw (Amondawa), ait (Arikem), ama (Amanayé), api (Apiaká), aqz (Akuntsu), arr (Arara-Karo), arx (Aruá), asn (Xingú Asuriní), asu (Tocantins Asurini), aux (Aurá), avv (Avá-Canoeiro), awe (Awetí), awt (Araweté), cin (Cinta Larga), cod (Cocama), eme (Emerillon), gn, gn-cls (Classical Guarani), gnw (Western Bolivian Guarani), gub (Guajajára), gug (Paraguayan Guarani), gui (Eastern Bolivian Guarani), gun (Mbya Guarani), guq (Aché), gvj (Guajá), gvo (Gavião do Jiparaná), gyr (Guarayu), jor (Jorá), jua (Júma), jur (Jurúna), kay (Kamayurá), kgk (Kaiwá), kpn (Kepkiriwát), ktn (Karitiâna), kuq (Karipuna), kyr (Kuruáya), kyz (Kayabí), mav (Sateré-Mawé), mdz (Suruí Do Pará), mnd (Mondé), mpu (Makuráp), msp (Maritsauá), myu (Mundurukú), nhd (Chiripá), omg (Omagua), oym (Wayampi), paf (Paranawát), pah (Tenharim), pak (Parakanã), pog (Potiguára), psm (Pauserna), pta (Pai Tavytera), pto (Zo'é), pur (Puruborá), skf (Mekéns), srq (Sirionó), sru (Suruí), taf (Tapirapé), tkf (Tukumanféd), tpj (Tapieté), tpk (Tupinikin), tpn (Tupinambá), tpr (Tuparí), tpw (Old Tupi), tqb (Tembé), tup-gua, tup-gua-pro (Proto-Tupi-Guarani), tup-kab (Kabishiana), tup-pro (Proto-Tupian), twt (Turiwára), urb (Urubú-Kaapor), urp (Uru-Pa-In), uru (Urumi), urz (Uru-Eu-Wau-Wau), wir (Wiraféd), wyr (Wayoró), xaj (Ararandewára), xet (Xetá), xiy (Xipaya), xmo (Morerebi), yrl (Nheengatu), yuq (Yuqui)
|
zlw-lch (Lechitic) |
csb (Kashubian), pl (Polish), pox (Polabian), szl (Silesian), zlw-opl (Old Polish), zlw-pom, zlw-slv (Slovincian)
|
The list is maintained in Module:etymon/data/text_allowed. Mode is warn (off = disabled, warn = warning only, error = enforce).
Language-specific behavior
Some languages have special handling configured in the module:
Chinese
Etymology trees and text are disabled for Chinese languages (those in the zhx family, excluding contact languages (e.g. creoles, pidgins, mixed languages)). See Wiktionary:Beer parlour/2025/May#Template:etymon for Chinese for discussion. Categories are also suppressed, and transliterations are not displayed.
Finnish
The :af keyword is non-transitive for Finnish, meaning the etymology chain stops at affixed terms.
Examples
(on English father)
{{etymon|en|:inh|enm:fader<id:father>|id=male parent}}
This means: father is inherited from Middle English fader.
(on a page with only one etymology)
{{etymon|en|:inh|enm:fader}}
This also works when there is only one etymology—no ID required on either side.
(on Proto-Indo-European *ph₂tḗr)
{{etymon|ine-pro|:af|*peh₂-<id:protect><unc>|*-tḗr<id:agent noun><unc>|id=father}}
This means: *ph₂tḗr might come from *peh₂ + *-tḗr. In this case, the etymons are associated with the derivation keyword ":af". Note that since the language is not specified for either etymon, the template assumes that the two etymons are ine-pro (Proto-Indo-European).
(on Polish podłoga)
{{etymon|pl|:deverbal|podłożyć<id:put>}}
This means: podłoga comes from Polish podłożyć.
{{etymon|en|:bor|fr:bouquet<id:bundle><ref:{{R:TLFi}}>|id=bundle of flowers|text=+}}
This means: bouquet is borrowed from French bouquet, with a reference to the TLFi dictionary. The reference will appear after the period in the generated text: "Borrowed from French bouquet.[1]"
{{etymon|en|:af|un-|happy<aftype:naf>|id=example}}
Using <aftype:naf> to mark "happy" as a non-affix (the base word) rather than letting auto-detection determine the affix type.
{{etymon|en|:root|ine-pro:*bʰer-<id:carry>|id=example}}
Using :root to categorize the term under "English terms derived from the Proto-Indo-European root *bʰer-" without showing this in the tree or text.
{{etymon|en|:der|la:exemplum<ety:inh<la:eximō>>|id=example}}
Using inline etymology to specify that Latin exemplum is inherited from eximō, useful when the Latin page doesn't have an etymon template.
{{etymon|en|:bor<conj:and/or>|es:cruzado|pt:cruzado|text=+}}
Using the <conj:and/or> modifier to specify a custom conjunction. This produces: "Borrowed from Spanish cruzado and/or Portuguese cruzado."
Suppressed terms
{{etymon|en|:inh|enm:-<t:some term>|id=example}}
Using - as the term to suppress the term display and show only the language. The gloss and other modifiers are still displayed.
Text output modes
{{etymon|en|:inh|enm:word|id=example|text=+}}
Using |text=+ to show only the immediate etymon (one step).
{{etymon|en|:inh|enm:word|id=example|text=++}}
Using |text=++ to show the full etymology chain (all steps, following through to the ultimate origin).
{{etymon|en|:inh|enm:word|id=example|text=*}}
Using |text=* to show the etymology chain until reaching the first existing entry (blue link).
{{etymon|en|:inh|enm:word|id=example|text=3}}
Using |text=3 to show up to 3 steps in the etymology chain.
{{etymon|en|:der|ar:كلمة|id=example|text=:ar}}
Using |text=:ar to show the etymology chain until reaching an Arabic term. Invalid language codes will trigger a warning and default to showing the full chain.
{{etymon|en|:bor|ota:پاشا<tr:paşa><ety:bor<fa-cls:پَادْشَاه>>|id=title|text=:ota*}}
Using |text=:ota* on English pasha to stop at Ottoman Turkish if it exists, otherwise continue to Persian.
Multiple etymon sources
{{etymon|en|:bor<unc>|es:palabra|pt:palavra|id=example}}
Multiple etymons under the same keyword with uncertainty. The <unc> modifier on the keyword marks the entire derivation as uncertain ("Possibly borrowed from...").
{{etymon|en|:inh|enm:word|:bor|fr:mot|id=example}}
Multiple keywords in sequence. This indicates the term is both inherited from Middle English word and also borrowed from French mot (e.g., for different senses or competing etymologies).
{{etymon|en|:af|un-<aftype:pre>|happy<aftype:naf>|-ness<aftype:suf>|id=example}}
Complex affixation with explicit affix types. The <aftype:...> modifier specifies whether each component is a prefix (pre), suffix (suf), infix (in), interfix (inter), circumfix (circum), non-affix base word (naf), or root (root).
Semantic relationships
{{etymon|en|:influence|en:comfort<id:ease>|id=modern sense}}
Using :influence to indicate semantic influence. This is non-transitive (does not continue the etymology chain). Useful for cases like discomfit being influenced by discomfort.
{{etymon|en|:calque|de:Weltanschauung|id=example}}
Using :calque for calques. Non-transitive; the etymology of the source term is not followed.
{{etymon|en|:pcal|fr:gratte-ciel|id=example}}
Using :pcal (partial calque). Also non-transitive.
{{etymon|en|:sl|la:musculus<id:muscle>|id=example}}
Using :sl (semantic loan). Also non-transitive.
Specialized borrowings
{{etymon|en|:lbor|la:exemplum|id=example}}
Using :lbor for learned borrowings (terms borrowed through literary or scholarly transmission rather than natural language contact).
{{etymon|es|:slbor|la:capitulum|id=example}}
Using :slbor for semi-learned borrowings (learned borrowings that have been partially adapted to native phonology).
{{etymon|en|:ubor|fr:déjà vu|id=example}}
Using :ubor for unadapted borrowings (terms borrowed with original orthography intact).
{{etymon|ja|:obor|zh:電話|id=example}}
Using :obor for orthographic borrowings (borrowing of written characters/script).
Word formation
{{etymon|en|:blend|motor|hotel|id=example}}
Using :blend for blends (portmanteau words).
{{etymon|en|:univ|good|bye|id=example}}
Using :univerbation (or :univ) for univerbation (multiple words becoming one).
{{etymon|en|:clip|advertisement|id=example}}
Using :clipping (or :clip) for clippings.
{{etymon|en|:bf|editor|id=example}}
Using :back-formation (or :bf) for back-formations.
{{etymon|en|:apheretic|esquire|id=example}}
Using :apheretic (or :apheresis) for apheretic forms (loss of initial unstressed vowel).
{{etymon|en|:denominal|hammer|id=example}}
Using :denominal (or :denom) for denominal verbs (verbs derived from nouns).
{{etymon|en|:deverbal|speak|id=example}}
Using :deverbal for deverbals (nouns/adjectives derived from verbs).
Inline etymology for redlinks
{{etymon|en|:der|xno:mot<ety:inh<fro:mot<ety:inh<la:muttum>>>>|id=example}}
Using nested inline etymologies to specify a full chain for terms without entries. This specifies that Anglo-Norman mot is inherited from Old French mot, which is inherited from Latin muttum.
{{etymon|en|:der|enm:word<ety:af<ang:word><-s>>|id=example}}
Inline etymology with affixation—specifies the Middle English term is formed from Old English word + -s.
References
{{etymon|en|:bor|fr:mot<ref:{{R:TLFi}} !!! {{R:Larousse}}>|id=example}}
Multiple references on a single etymon, separated by !!! (with spaces).
{{etymon|en|:bor<ref:{{R:OED}}>|fr:mot|id=example}}
Reference on the keyword itself (applies to the derivation relationship, not just the term).
{{etymon|en|:der|la:verbum<ref:<<name:LS>>{{R:L&S}}>|id=example}}
Named reference using <<name:LS>> syntax, which allows re-use elsewhere with <ref:<<name:LS>>>.
{{etymon|en|:der|la:verbum<ref:<<group:etymology>>{{R:L&S}}>|id=example}}
Reference with a group, for use with grouped reference lists.
Tree display
{{etymon|en|:inh|enm:word|id=example|tree=1}}
Using |tree=1 to display an etymology tree.
Combined features
{{etymon|en|:inh|enm:payen<id:pay><ref:{{R:MED}}>|id=give money|text=++|tree=1}}
Comprehensive example combining: etymology ID, inheritance keyword, etymon with ID and reference, full text output, and tree display.
{{etymon|en|:bor<unc>|fr:exemple<t:example><tr:ɛɡzɑ̃pl>|etydate=1500s<ref:{{R:OED}}>|id=example|nodot=1|text=+}}
Example with attestation date, uncertain borrowing, gloss and transliteration, single-step text output, and no final period (for combining with additional text).
See Template:etymon/testcases for test cases and more examples of use.
Categorization
The template generates various categories depending on how it is used, including the ones within:
- Category:Entries referencing ambiguous etymons by language
- Category:Entries referencing etymons with mismatched IDs by language
- Category:Entries with etymology trees by language
- Category:Entries with etymology texts by language
- Category:Pages with etymology trees
- Category:Pages with inline etymon for redlinks
- Category:Pages with redundant inline etymon
- Category:Pages using etymon with no ID
- Category:Terms coined ex nihilo by language
as well as various standard etymology categories depending on what entries exist in the tree:
- Borrowing categories (e.g., Category:English terms borrowed from French)
- Inheritance categories (e.g., Category:English terms inherited from Middle English)
- Affix categories (e.g., Category:English terms prefixed with un-)
- Root categories (e.g., Category:English terms derived from the Proto-Indo-European root *bʰer-)
- Specialized borrowing categories (learned, semi-learned, orthographic, unadapted)
- Calque, partial calque, and semantic loan categories
Tracking
The module tracks various statistics for debugging and analysis:
- Tree depth (how many generations deep the etymology goes)
- Number of nodes (total terms in the tree)
- Number of unique languages
- Linear vs. branching trees
These can be found in the tracking categories under Template:tracking/etymon/.
References
TemplateData
TemplateData for etymon
This template may be used indicate a term's immediate ancestor(s).
| Parameter | Description | Type | Status | |
|---|---|---|---|---|
| Language | 1 | The language of the current entry.
| String | required |
| Etymology ID | id | The ID of the current etymology section. Required when there are multiple etymologies for the same term.
| String | suggested |
| Title | title | The title of the current entry (if different from the page title)
| String | optional |
| Part of speech | pos | The part of speech (prefix, suffix, interfix, infix, root, word)
| String | optional |
| Request for etymology | rfe | Adds a request for etymology notice. Set to 1 for standard notice, or provide custom text. Supports inline modifiers.
| String | optional |
| Etymon parameter #1 | 2 | The first etymon parameter. Typically for derivation keyword.
| String | optional |
| Etymon parameter #2 | 3 | The second etymon parameter. | String | optional |
| Etymon parameter #3 | 4 | The third etymon parameter. | String | optional |
| Etymon parameter #4 | 5 | The fourth etymon parameter. | String | optional |
| Etymon parameter #5 | 6 | The fifth etymon parameter. | String | optional |
| Etymon parameter #6 | 7 | The sixth etymon parameter. | String | optional |
| Etymon parameter #7 | 8 | The seventh etymon parameter. | String | optional |
| Etymon parameter #8 | 9 | The eighth etymon parameter. | String | optional |
| Etymology date | etydate | The date or time period when the term was first attested
| String | optional |
| Ex nihilo | exnihilo | Set to categorize as a term coined ex nihilo | Boolean | optional |
| Text | text | Automatically generates some text. Values: + (single step), ++ (all steps), * (to nearest blue link), a number (max steps), :langcode (stop at language), :langcode* (stop at language or first blue link if redlink)
| String | optional |
| Tree | tree | Set this to "1" to display a tree.
| String | optional |
| Non-lemma | nl | Set to consider the entry a non-lemma | Boolean | optional |
| No categories | nocat | Set to suppress automatic categorization | Boolean | optional |
| No dot | nodot | Set to remove the final period from generated text | Boolean | optional |