A person gets "62% French & German" from 23andMe, then obtains Global25 coordinates, runs an era-scoped fit, and finds nothing called French or German anywhere — instead proportions of early farmers, steppe pastoralists and foragers, or of Iron Age and medieval populations. The instinctive question is which one is wrong. The correct answer is that they are not answers to the same question, and each is routinely misread as the other.
What a testing-company estimate computes#
23andMe's Ancestry Composition, AncestryDNA's ethnicity estimate and their peers work segment by segment: each stretch of your chromosomes is assigned to the modern reference population it best matches, out of a proprietary panel of living, recently-rooted people; the percentages are the totals, smoothed and calibrated. Three properties follow. The references are modern — the categories are shaped like today's countries and labelled accordingly. The time horizon is shallow — the estimate approximates where your ancestors of the last few centuries would plausibly be filed. And the pipeline is closed — panels, priors and smoothing are unpublished and change between releases, which is why estimates jump on update days without anyone's DNA changing.
None of that makes it bad. For its own question — which present-day populations do the segments of my genome file under — a big proprietary panel is genuinely strong, and updates usually make it stronger. The failure mode is only the label: "French & German" names a reference bucket, not a documented ancestor.
What a G25 model computes#
A G25 analysis is open arithmetic over published references: your row, a stated panel of population averages (ancient or modern, era by era), a fit distance anyone can re-run. The time horizon is whatever the panel is — Bronze Age sources give formation-era proportions, modern-era panels give present-day resemblance, and the distance ranking gives the raw neighbourhood with no model at all.
The trade is symmetrical. The G25 stack is transparent, reproducible and time-scoped — and it is built on chip-scale markers, one fixed projection, and reference panels a fraction the size of a testing company's, with no test attached to any fit. The corporate estimate is opaque and shallow — and segment-level, hugely sampled, and calibrated against customers with known ancestry at a scale nobody else has.
Why specific disagreements happen#
- Different clock. "62% French & German" and "48% Anatolian farmer" can both be right: one describes the last ~300 years, the other the last ~8,000. Most confusion is just this.
- Different buckets. Corporate categories follow modern borders; G25 panels follow sampled populations. A Balkan genome gets filed under whichever national buckets the company trained, while a G25 era fit reads it as the layered history those buckets share.
- Trace components. Corporate estimates smooth small signals with priors (and still print phantom traces); coordinate fits round noise onto available sources. Different machinery, same rule: sub-few-percent figures are unconfirmed in both.
- Update whiplash. Your estimate changed; your row cannot. A fixed coordinate re-analysed against growing reference panels is the opposite failure mode of a fixed genome re-filed by a changing algorithm. Knowing which one moved tells you what the "change" means.
Which to use for which claim#
| The claim | The right instrument |
|---|---|
| "Where were my recent ancestors likely from?" | The testing company's estimate (and its match list — matching beats admixture here) |
| "Which populations do I resemble, today and in the past?" | G25 distances, era by era |
| "What deep ancestries formed my genome, in what proportions?" | An era-scoped G25 fit; the worked version is the Global25 analysis |
| "Is this component real / is this source required?" | Neither — that is a test, and coordinate fits and corporate estimates both lack one. qpAdm is the method with a p-value. |
The two systems make each other more readable, not less. The corporate estimate is a well-funded answer about the recent past; the G25 stack is an open answer about the deep past — and an ancient-DNA analysis is what the raw file you already paid for is still capable of telling you.
Terms used here are defined in the glossary.



