Tajik DNA
Genetic origins and closest populations · Tajikistan and Afghanistan · Central Asia
Tajiks speak Tajik, a form of Persian, which makes them the main Iranian-speaking people of Central Asia among mostly Turkic-speaking neighbours. They live in Tajikistan, Afghanistan and the old cities of Uzbekistan, and the Samanid dynasty of the 9th and 10th centuries, centred on Bukhara, is remembered as a golden age of Persian literature. In our model, the Tajik average is 31.2% steppe, 21.4% Iranian Neolithic, 16.8% Anatolian Neolithic farmer, 10.9% East Asian, 9.6% Ancient Northeast Asian, 7.3% South Asian proxy and 2.9% Arctic Siberian. Its Ancient Northeast Asian share is far below that of the Uzbeks (22.3%) or Kazakhs (47.6%), while the mountain Pamiri and Yaghnobi carry even more steppe ancestry. Language does not wall off the neighbours, though: after other Tajik samples from Samarkand and Tashkent, the closest are the Turkic-speaking Turkmen (0.031) and Uzbeks from Afghanistan. The closest ancient groups are only moderately close, led by Early to Late Modern nomads of the Tian Shan (0.044).
- Largest ancient components in our model: Steppe herders (Yamnaya) 31.2%, Iranian Neolithic farmers 21.4%, Anatolian Neolithic farmers 16.8%
- Closest other population in our data: Turkmen, distance 0.0310
- Closest ancient group in our data: Kyrgyzstan Early-Late Modern Nomad Tian Shan, distance 0.0440
Where does Tajik DNA sit on the genetic map of the world? To answer, we take the averaged Global25 (G25) coordinates of the Tajik individuals sampled in Tajikistan and Afghanistan and compare them with deep ancestral reference populations, with other modern groups and with ancient genomes. It is one of 335 populations in our atlas.
Ancient make-up of the Tajik average
The ancient make-up of the Tajik average is led by Steppe herders (Yamnaya) (31.2%), followed by Iranian Neolithic farmers (21.4%) and Anatolian Neolithic farmers (16.8%). The Steppe herders (Yamnaya) component reflects the Bronze Age herders of the Pontic-Caspian steppe (Yamnaya and related groups), whose ancestry spread across Europe and into Asia from about 3000 BC. The Iranian Neolithic farmers share (21.4%) reflects the early farmers and herders of the Zagros mountains in Iran. No single source dominates, so the profile is best read as a balanced blend of several ancient ancestries.
Smaller traces of Arctic Siberian (Nganasan-related) (under 5%) also appear. At that level they can reflect real minor ancestry, but also simple model noise, so they should not be over-read. The model fits well (fit distance 0.018), so these proportions are a reasonable summary. This population was modelled with a global set of reference populations (ancient genomes, plus modern stand-ins where no suitable ancient genome exists), because West Eurasian sources alone cannot describe it.
Simorgh – Iranian Plateau Report
Closest modern populations to Tajik
Among the modern populations in the G25 data, the closest to the Tajik average is Tajik (Samarkand), at a distance of 0.026 (close). Next come Tajik (Tashkent) at 0.028 and Turkmen at 0.031. Leaving aside the other regional samples listed under the same name in the G25 sheet, the nearest other populations are Turkmen (0.031), Uzbek (Afghanistan) (0.034) and Turkmen (Uzbekistan) (0.034). By the tenth closest (Sarikoli (China)) the distance grows to 0.065, a common pattern for a population that shares ancestry with its neighbours while keeping a profile of its own.
| # | Population | Distance | Closeness |
|---|---|---|---|
| 1 | Tajik (Samarkand) | 0.0256 | Close |
| 2 | Tajik (Tashkent) | 0.0284 | Close |
| 3 | Turkmen | 0.0310 | Close |
| 4 | Uzbek (Afghanistan) | 0.0341 | Close |
| 5 | Turkmen (Uzbekistan) | 0.0343 | Close |
| 6 | Turkmen (Iran) | 0.0425 | Close |
| 7 | Uzbek (Tashkent) | 0.0552 | Moderate |
| 8 | Pamiri (Sarikoli) | 0.0580 | Moderate |
| 9 | Turkmen (Golestan) | 0.0614 | Moderate |
| 10 | Sarikoli (China) | 0.0647 | Moderate |
Distances are Euclidean distances between averaged G25 coordinates. On our scale, below 0.025 is very close, below 0.050 close, below 0.080 moderate, and beyond that distant.
Closest ancient populations to Tajik
If we search ancient DNA for the best match to Tajik, Kyrgyzstan Early-Late Modern Nomad Tian Shan (Modern era, c. 1500-1950 AD) comes first, at 0.044, with Kazakhstan High Medieval Kipchak Nurataldy next. That is a close match: much of the modern profile was already present then, while later movements of people added further layers. A close ancient match is not proof of direct descent: it means those individuals carried a similar overall mix of ancestry.
| # | Ancient sample or group | Period | Distance |
|---|---|---|---|
| 1 | Kyrgyzstan Early-Late Modern Nomad Tian Shan | Modern era, c. 1500-1950 AD | 0.0440 |
| 2 | Kazakhstan High Medieval Kipchak Nurataldy | High Medieval, c. 1000-1300 AD | 0.0488 |
| 3 | Xinjiang HP Junmachanyilian (South Central Asian Profile) | Historical period | 0.0582 |
| 4 | Xinjiang HP Hetian | Historical period | 0.0599 |
| 5 | Mongolia Early Medieval Turkic Period (South Central Asian Profile) | Early Medieval, c. 500-1000 AD | 0.0649 |
| 6 | Turkey Anatolia Early Modern Capalibag (Turkic Profile) | Early Modern, c. 1500-1800 AD | 0.0656 |
| 7 | Ukraine Early Medieval Saltovo-Mayaki Bochkove (Bulgar, East Asian-Mixed Profile) | Early Medieval, c. 500-1000 AD | 0.0661 |
| 8 | Kazakhstan LBA Talapty | Late Bronze Age | 0.0671 |
| 9 | Kyrgyzstan IA Saka Chilpek | Iron Age | 0.0675 |
| 10 | Kyrgyzstan IA Uch-Kurbu | Iron Age | 0.0689 |
Ancient DNA from Tajikistan and Afghanistan
Our ancient DNA database holds 14 individuals excavated in present-day Tajikistan and Afghanistan, dated from about 6,273 BC to 8 BC. The most frequent Y-DNA haplogroups among them are R2a (2), E1b (1) and G2a (1), and the most frequent mtDNA haplogroups are U2 (4), J1 (2) and T2 (2). These are people who lived on the same land in the past, not necessarily ancestors of today's Tajik population.
Most frequent Y-DNA haplogroups
Browse them in our ancient DNA database: 13 from Tajikistan, 1 from Afghanistan.
Compare yourself with Tajik
Paste your G25 coordinates (scaled, one line, with or without a name in front) and we compute your genetic distance to the Tajik average and to its closest neighbours, right in your browser. Nothing is uploaded or stored.
| # | Population | Your distance | Closeness |
|---|
This list only covers Tajik and its neighbours. To find out which of our reports actually fits your DNA, run the free Report Finder: it runs the same fit test that every report uses before an order.
No G25 coordinates yet? Get a free simulated G25 from your raw DNA file, or order G25 coordinates.
About these numbers
Keep in mind that a population average is a statistical summary of a limited number of sampled individuals. Real people inside any group vary, some carry more of one ancestry and some less, and no genetic profile decides who is or is not Tajik. Use these numbers as a map of deep ancestry, not as a label.
Method: averaged G25 coordinates, Euclidean distances to other modern and ancient averages, and a non-negative least-squares model against deep ancestral reference populations (data generated 2026-10-01). Read the full method.
Go deeper than the average
This page describes the Tajik average. Your own DNA has its own story: Simorgh – Iranian Plateau Report (15€) models your genome against the ancient sources and modern communities of the region, era by era, in a personal PDF report.
A report built for people of Iranian-plateau ancestry: ancient regional core (Bronze/Iron Age Iranian Plateau, Achaemenid, Parthian, Sassanid Persia) versus non-regional outside admixture, across 10 communities (Persian, Lur/Bakhtiari...
Frequently asked questions
What is the ancient genetic make-up of Tajik?
Modelled with deep ancestral reference populations, the Tajik average is about 31.2% Steppe herders (Yamnaya), 21.4% Iranian Neolithic farmers and 16.8% Anatolian Neolithic farmers. These proportions are model estimates for a group average, not exact values for any one person.
Which populations are genetically closest to Tajik?
Leaving aside other regional samples listed under the same name in the G25 sheet (such as Tajik (Samarkand)), the closest modern populations to the Tajik average are Turkmen, Uzbek (Afghanistan) and Turkmen (Uzbekistan). The closest ancient matches are Kyrgyzstan Early-Late Modern Nomad Tian Shan, Kazakhstan High Medieval Kipchak Nurataldy and Xinjiang HP Junmachanyilian (South Central Asian Profile).
Can a DNA test tell me if I am Tajik?
No DNA test can confirm an ethnicity or a nationality. What DNA can show is how similar your genome is to the sampled Tajik average and to its neighbours. If you have G25 coordinates, paste them in the comparison box on this page to check that for free.
Which ExploreYourDNA report suits Tajik ancestry?
Simorgh – Iranian Plateau Report is our regional report for this part of the world. On the Tajik average its fit test reads partial, so check your own DNA first: the free Report Finder runs that test on your file.
Where does the data on this page come from?
From the averaged Global25 (G25) coordinates of the sampled individuals: genetic distances to other modern and ancient population averages, and a non-negative least-squares model against deep ancestral reference populations. The full method is described at https://www.exploreyourdna.com/populations#method.