Skip to content
Genetics

The DNA of the Uzbeks: population genetics of Central Asia

Uzbeks sit genetically at the intersection of Western Eurasian and East Eurasian ancestries — a signature of thousands of years of Silk Road movement, Iranian farming, and Turkic-Mongol expansion.

Edited by Reviewed by Last reviewed on

9 min read5 sourcesPeople & Language

Modern Uzbeks are one of the most-studied populations in Central Asian genetics precisely because they lie at a hinge. Uzbekistan sits between the Iranian farming world to the west and the Turkic-Mongol steppe world to the east, and its genetic profile records both. This article summarises what population geneticists have actually found, with proper citations to the primary literature.

How to think about Central Asian ancestry

Central Asia has been a corridor of human movement for at least the last five thousand years. In roughly chronological order it has absorbed: Iranian-speaking farmers (Bactria-Margiana Archaeological Complex, ~2300 BCE); Iranian-speaking Sogdians and Scythians (1st millennium BCE); Hellenistic Greeks (from 329 BCE, briefly); Kushans and Hephthalites (1st–6th centuries CE); Arab conquerors (7th–8th centuries); and Turkic and Mongol steppe expansions (from the 6th century, peaking in the Turkic Kaganate and the 13th-century Mongol conquest). Every one of those movements left a trace in modern DNA.

The most important single frame for reading Uzbek genetics is that Uzbekistan sits genetically between Western Eurasia (Europe, Iran, the Caucasus) and Eastern Eurasia (Mongolia, north-east Asia). Autosomal (genome-wide) studies typically put Uzbeks around 60–75 per cent West Eurasian and 25–40 per cent East Eurasian — a proportion that shifts by region within Uzbekistan itself.

What autosomal (genome-wide) studies show

Martínez-Cruz et al. (2011) analysed 26 autosomal short tandem repeats and 15 Y-chromosome markers in nine Central Asian populations. Their principal finding is that Central Asian genetic diversity is structured primarily by language family: Turkic-speaking populations (including Uzbeks) form a cluster more closely related to East Asian populations, while Indo-Iranian-speaking populations (Tajiks) cluster closer to Western Eurasians. Uzbeks show an intermediate profile — as one would expect from a Turkic-speaking population that historically settled on top of an Iranian-speaking substrate.

Yunusbayev et al. (2015) — the largest study to date of Turkic-speaking populations — used genome-wide SNP data to trace the impact of the medieval Turkic expansions. They found that Turkic peoples across Eurasia (from Anatolia through Central Asia to Siberia) share signals of ancestry from an ancestral homeland in south Siberia and Mongolia. For Uzbeks, this component is present but overlaid on a much larger substrate of local (Iranian-farmer-descended) ancestry.

In simple terms: modern Uzbeks are mostly descended from the Iranian-speaking oasis farmers who lived here for millennia, with a significant genetic infusion from the medieval Turkic and Mongol movements. Language changed faster than genes.

Mitochondrial DNA (maternal lineages)

Irwin et al. (2010) sequenced the mtDNA of over 1,500 individuals across Uzbekistan — Fergana, Karakalpakstan, Khorezm, Kashkadarya, and Tashkent — plus samples from neighbouring countries. The Uzbek mtDNA composition is roughly 65–70 per cent West Eurasian haplogroups (H, U, T, K, J, HV) and 30–35 per cent East Eurasian (D, C, F, M, G). The exact proportions shift by region.

Regional patterns match history. Karakalpakstan (Kipchak Turkic, closer to Kazakh) shows more East Eurasian mtDNA. Kashkadarya and Bukhara show more West Eurasian mtDNA (consistent with older Sogdian/Iranian substrate). Fergana Valley — where Turkic settlement was heaviest — shows intermediate values.

Y-chromosome (paternal lineages)

Y-DNA studies (Karafet et al., Wells et al., Underhill et al.) find several dominant haplogroups in Uzbek men:

  • R1a (~25–35 %) — associated with Bronze Age Indo-European expansions, especially the Andronovo culture. High in the Fergana Valley.
  • J2 (~10–20 %) — associated with the Iranian farming world and the early Silk Road.
  • C2 / C-M217 (~10–20 %) — associated with East Asian and Mongolic paternal lineages, including a specific branch traditionally linked to Genghis Khan and his descendants.
  • Q (~5–10 %) — an old northern-Eurasian lineage.
  • G (~5 %) — associated with the Caucasus and early Anatolian farmers.
  • R1b, N, O — smaller shares.

What this actually means

Three big-picture conclusions can be drawn safely from the literature:

First, Uzbeks are not simply Turkic incomers. The autosomal, mitochondrial, and Y-chromosome data all point to a population that is majority-descended from the Iranian-speaking farmers who lived here from the Bronze Age onwards, with a real but proportionally smaller Turkic-Mongol overlay.

Second, Turkic-speaking populations of Central Asia share a common ancestral signal, consistent with a real medieval expansion from a homeland in south Siberia / Mongolia. This shows up in Uzbeks too.

Third, regional variation within Uzbekistan matches settlement history. Fergana Valley and Karakalpakstan carry more steppe ancestry; Bukhara and Samarkand carry more Iranian substrate.

A note on consumer DNA tests

Consumer DNA tests (23andMe, AncestryDNA, MyHeritage) tend to compress Uzbek ancestry into categories like 'Central Asian', 'Uzbek/Turkmen', 'Mongolian & Manchurian', or a combination. This is roughly consistent with the academic picture but coarser: the actual archaeological and linguistic complexity of Uzbek ancestry is deeper than a pie chart. For a genuinely rigorous view, the primary literature cited below is the right place to look.

References

Every article on this site cites primary and secondary sources. Inline superscripts [n] link to the numbered entries below. All external links open in a new window.

  1. Yunusbayev et al., 2015 — The Genetic Legacy of the Expansion of Turkic-Speaking Nomads across Eurasia— PLOS Genetics (2015)
  2. Martínez-Cruz et al., 2011 — In the heartland of Eurasia: the multilocus genetic landscape of Central Asian populations— European Journal of Human Genetics (2011)
  3. Irwin et al., 2010 — The mtDNA composition of Uzbekistan: a microcosm of Central Asian patterns— International Journal of Legal Medicine (2010)
  4. Karafet et al., 2002 — High levels of Y-chromosome differentiation among native Siberian populations and the genetic signature of a boreal hunter-gatherer way of life— Human Biology (2002)
  5. Narasimhan et al., 2019 — The formation of human populations in South and Central Asia— Science (2019)
Continue reading

Related articles