Start here
A beginner’s roadmap: from one raw DNA file to a family tree with your ancestors’ DNA filled in. You do not need to do this in one sitting — but the steps are in the order that works.
The big picture
Everything here serves one of two goals:
- Build the tree — people, relationships, dates and places. Every tested kit gets attached to its person, so the tree and the DNA point at each other.
- Fill in the ancestral DNA — when several matches all share the same piece of a chromosome, that piece can be accumulated onto the ancestor who passed it down. Those inferred pieces are called Tree DNA, and the Chromosome Map draws them per chromosome, coloured by ancestor.
The two goals feed each other in one loop:
↻ then repeat — every new relative sharpens the map
You will still use the other guides for the detail; this page is the map. If a term below is unfamiliar, the jargon buster at the bottom has one-line definitions.
1Get your DNA in
Goal: your raw DNA — and as many relatives’ kits as you can gather — uploaded and being compared.
- Create a free account and confirm your email.
- Download the raw data file from the company that tested you (see Downloading your file).
- Upload it exactly as downloaded —
.txt,.csv,.zipor.gz. Your kit gets a kit number and is compared against every other kit. - Repeat for any relative who has tested. Parents, siblings, aunts, uncles and cousins are each worth far more than another ethnicity estimate.
Why relatives matter: everything later — triangulation, phasing, reconstruction — gets easier with more tested people. Three matches on one segment is what turns a guess into evidence; a tested parent or sibling is what lets you separate your two sides.
2Meet your matches
Goal: a shortlist of your closest and most promising matches, and a feel for how cM and segments behave.
- Open One-to-Many and read down the list: shared cM, segments and largest segment. The top of the list is your close family.
- Click Review on a match for the full match page — relationship estimate, shared matches and a chromosome browser.
- Click Compare (or open One-to-One) to see the exact segments you share.
- Start labelling as you go: Tag Groups are your own private colour labels, and each match card takes a private note.
Reading the result: one big segment is stronger evidence than many tiny ones; segments under about 7 cM usually are not evidence of a real relationship. The matching guide explains cM, IBD1/IBD2 and the relationship ranges in depth.
3Check for endogamy
Goal: know how much to trust the cM numbers before you build on them.
Run Are Your Parents Related? once, early. It scans your kit for long runs of homozygosity (ROH) — stretches where both copies of a chromosome carry the same DNA.
- No significant ROH → the standard relationship ranges apply fairly directly.
- Long ROH → your parents share an ancestor, and so probably do many of your matches. Distant relatives will share more cM, in more segments, than the standard tables assume.
In Roma and other endogamous communities this is common and expected — it is community history, not a scandal. The practical takeaway: lean on the largest segment, triangulation and the tree rather than the cM total alone. See the endogamy notes in the matching guide.
4Split matches into family lines
Goal: turn a flat match list into groups that look like ancestral lines — then prove a group with segments.
- The Leeds method groups your matches into (up to) four grandparent groups — the classic first way to split a match list into family lines.
- Identify one known relative per group and the rest of the group is probably on that person’s branch. Label it with a Tag Group so the match list stays readable.
- Segment Search lists every shared segment, so you can find everyone who shares a particular region.
- Triangulation takes it further: it finds clusters of matches who all share the same segment with each other. That is the point where a segment becomes evidence of one common ancestor — and the tool can add it to your tree (step 7).
Hands off to: named or hypothesised lines, and triangulated segments ready to be painted onto ancestors.
5Build the tree and attach kits
Goal: a tree that contains both the people and the tested DNA, so every later tool has somewhere to put its result.
- Open the tree viewer. Every registered user can edit the shared tree; right-click a card for the menu (edit, add relative, set home). It is modelled on the way Ancestry’s family tree works, though nowhere near as sophisticated — it should do the job for now.
- Add the people you know — parents, grandparents, siblings, the aunts and uncles — and attach each tested kit to its person (from DNA Kits → kit settings → tree linking, or from the person’s page).
- Set your own home person with the star, so the tree opens where you want.
- DNA badges are on by default — every tested person who matches the tree’s selected DNA test carries a green DNA {cM} badge, and inferred Tree DNA people carry a blue one. The bottom rail lists that test’s top DNA matches: click a name to jump to them in the tree, or a dashed one to open its Review page. Then open Compare DNA to tree to see which branch a kit’s matches fit.
- If the same person appears twice, use the Review & merge banner on the person page.
Hands off to: a tree with tested people wired in — which is what makes step 7 possible. The tree guide covers badges, ThruLines and merging in detail.
6Place the matches that do not fit
Goal: find where a mystery match belongs instead of guessing.
- WATO (“What Are The Odds?”) scores every plausible place a match could sit in your tree. It compares the observed shared cM against what each placement would predict, using your real tested relatives as references, and ranks the placements.
- Read the top placements, then add the person to the tree under the best-supported one (or leave them floating until more evidence arrives). A “No possible combination found” message means no placement explains the cM — do not trust the top row.
- No tree at all yet? AutoKinship builds and scores possible arrangements from your matches alone.
Keep in mind: endogamy biases these estimates towards “too close”, and a predicted placement is a hypothesis until records or new testers confirm it. Neither tool is a self-learning model, and they do not both use the tree the same way: WATO is only as good as the tree — it scores against the tested relatives you have placed there, so every relative you add or placement you correct sharpens its next result — while AutoKinship works from your match list alone and does not depend on the tree at all. A thin or wrong tree gives WATO thin or wrong placements.
7Fill in your ancestors’ DNA the payoff
Goal: give the people in your tree the DNA they must have carried — even the ones who never tested.
The engine does this from evidence you have already collected. When a group of matches triangulates on one segment, the shared (ancestral) allele can be worked out, and it is accumulated onto the Tree DNA virtual kit of the ancestors the segment passed through. People who have a real kit never get inferred DNA — their real file is authoritative.
Pick the path that fits what you have:
| Your situation | Tool | What it produces |
|---|---|---|
| Several matches triangulate on the same segment (step 4) | Triangulation → Add all clusters to tree DNA | Inferred segments painted onto the Tree DNA of the ancestors in between |
| One parent tested, the other cannot be | Reconstruct a parent | A virtual kit (VK) for the untested parent, phased from a child against the known parent |
| An ancestor already has Tree DNA, and a child has a real kit | Reconstruct a parent (reverse-phase) | Alleles for the other parent — one more generation per run, re-run with a second child to accumulate more |
| Neither parent tested, but two or more siblings are | Visual phasing | Grandparent blocks you assign and paint onto the tree ancestors |
| Child and both parents tested | Phasing | Accurate maternal/paternal phased haplotypes |
| A deeper ancestor with several tested descendant lines | Lazarus | An ancestral kit accumulated from the descendants’ shared segments |
Then look at what you built:
- Tree DNA reconstructions — every virtual kit, with its SNPs, segments and coverage.
- The ancestor’s person page — the inferred segments assigned to them.
- Chromosome Map — your kit’s chromosomes drawn with each ancestor’s segments in their own colour (X follows inheritance).
- Compare DNA to tree DNA — run a real kit against the reconstructions with the matching engine.
- Tree conflicts — segments that could not be added because they disagree with what is already there. If the same block keeps being deleted by one person and added back by another, it is marked contested and can be escalated to an admin (from Visual Phasing): both versions cannot be right, so one of the branches is incorrect and needs checking against records. Escalated blocks and the full paint/remove history are listed on the conflicts page.
- A Tree DNA kit is a hypothesis, not a test result — never treat it as one. Coverage only grows as more descendants and segments are added; the rest is left blank, on purpose.
- A reconstruction built from descendants of both parents recovers the parent couple, not one individual, unless the descendants are branch-specific.
- Conflicting evidence is never silently overwritten — it is logged on the conflicts page for review.
- Do not compare two reconstructions to each other as if they were independent people.
8Keep the loop going
Goal: every new tester makes the map sharper — so keep feeding it.
- Invite relatives to test or upload. A new cousin can confirm a cluster, split a line, or extend a segment one generation further back.
- Re-run the grouping tools after new kits arrive: matching happens automatically, and clustering/triangulation recompute when you open them.
- Keep the tree and the DNA in sync — when a match is confirmed, attach them to their person; when a segment conflicts, review it.
- Side quests: Y-DNA and mtDNA follow your deep paternal and maternal lines; the Romani report and Ethnicity explore where your ancestry comes from. Those answer “where from” — this roadmap answers “who”.
How the tools fit together
| Tool | You give it | It gives you | Feeds into |
|---|---|---|---|
| 1 · Your DNA | |||
| Upload · DNA Kits | A raw data file | A kit number, a person node, automatic matching | Everything else |
| 2 · Matches | |||
| One-to-Many | A kit | Ranked matches with cM, segments, largest segment | Review, One-to-One, Clustering |
| Review | A match pair | Relationship estimate, shared matches, chromosome browser | One-to-One, tree |
| One-to-One | Two kits | Every shared segment + interactive chromosomes | Triangulation, Tree DNA |
| Are Your Parents Related? | A kit | ROH blocks / inbreeding estimate | How to read every cM figure |
| Segment Search | A kit, optionally a chromosome/range | All shared segments vs everyone | Triangulation |
| Triangulation | A kit + minimum cM | Clusters sharing the same segment | Tree DNA |
| 3 · Group & label | |||
| Leeds method | A kit + cM range | Four grandparent groups ≈ ancestral lines | Tag Groups, tree |
| Tag Groups | Matches or kits | Private labels and filters | Every match list |
| 4 · Tree | |||
| Tree viewer | People + relationships + kit links | The shared tree with DNA badges | All reconstruction tools |
| Compare DNA to tree | A kit | Which branch the matches fit | WATO |
| WATO · AutoKinship | A mystery match + reference kits | Ranked placements / scored trees | Tree edits |
| 5 · Fill in the ancestors | |||
| Triangulation → Add all to tree DNA | A triangulated cluster | Inferred segments on the ancestors in between | Tree DNA, Chromosome Map |
| Reconstruct a parent | A child + one known parent | A virtual kit for the missing parent | Tree DNA, Chromosome Map |
| Visual phasing · Phasing · Lazarus | Siblings, a trio, or descendants | Grandparent blocks / phased haplotypes / ancestral kits | Tree DNA |
| Tree DNA · Chromosome Map | Nothing new — it is already stored | The reconstructed kits and the ancestor-coloured map | Back to step 4, round again |
If you only do five things
- Upload your raw DNA.
- Open One-to-Many and review your five closest matches.
- Run Are Your Parents Related? before trusting any cM number.
- Group your matches (Leeds method) and tag the ones you recognise.
- Attach your closest tested relatives to their people in the tree.
Then, when three or more matches share the same segment, triangulate it and add it to Tree DNA. That is the moment your ancestors start to have DNA again.
Jargon buster
- Raw DNA
- The plain text file of your genotypes from the testing company — the file you upload, not the matches or ethnicity report built from it.
- SNP
- One position in the genome that a testing chip reads. Modern kits cover hundreds of thousands of them.
- cM (centimorgan)
- A unit of genetic distance shared with a match. More cM generally means a closer relationship — but endogamy inflates it.
- Segment
- One unbroken run of shared DNA, given as chromosome, start position and stop position.
- IBD1 / IBD2
- Shared on one chromosome copy (half-identical) / shared on both copies. Parent and child share everywhere at IBD1 with no IBD2; full siblings show both.
- Triangulation
- Three or more matches who all share the same segment with each other — strong evidence the segment came from one common ancestor.
- Phasing
- Working out which allele came from which parent, so a genotype can be split into a maternal and a paternal copy.
- Tree DNA / virtual kit
- A kit reconstructed for someone who never tested, built from descendants’ DNA. A hypothesis, not a test result; shown with a
VKkit number. - Endogamy
- Generations of marriage within one community, so distant relatives share more DNA than the standard relationship tables assume. Common in Roma families.
- ROH
- Runs of homozygosity — long stretches where both copies of a chromosome are identical, a sign that your parents share an ancestor.
- Kit number
- The short code for a kit:
AN,MH,LI,FT,M3for real uploads,VKfor Tree DNA,SKfor a superkit,LAZfor a Lazarus reconstruction.
Ready to start?
Upload your raw DNA, then come back to step 2 when your matches are ready.
Upload your raw DNA freeMethod notes: relationship ranges from the Shared cM Project 4.0 (Blaine T. Bettinger, CC 4.0); clustering follows Shared Clustering (Brecher); Y and mt haplogroups use yhaplo / ISOGG 2016 and PhyloTree Build 17 (van Oven, 2015); reference/IBD phasing follows Noto & Ruiz 2022 and uses the 1000 Genomes phase-3 panel; Global25 and the calculators are by Davidski. Tree DNA reconstructions are labelled as hypotheses throughout the site.