Getting Started With Family History Discovery
I picked up a DNA kit a few years back mostly because my cousin kept talking about how useful it was. The packaging looked like everything else on the market — clear plastic, a tube, instructions printed small enough to be annoying. I followed the steps, spat in the tube, mailed it off, and waited. Results came back in about six weeks. The raw data file alone was something like 700MB, which I didn't know what to do with at first. That's where most people get stuck. You have the results, but the actual work of making sense of them hasn't started yet. The initial ethnic estimate is the easy part. It's basically a statistical model comparing your DNA against reference populations the company has collected over the years. The numbers shift slightly between uploads and re-releases. That's normal. What people don't realize is that the ethnicity breakdown is the least useful thing in your results for doing actual genealogical research. The centimorgan matches and shared segments are what matter. That's the part you need to learn how to read. I used one of the free third-party upload tools to process the raw data after my first test. Gedmatch is the standard starting point. It lets you compare your matches across different databases and identify overlapping segments. Most of my early matches in the 200 to 400 centimorgan range turned out to be half-siblings or avuncular relationships that no one in the family had mentioned. That was the part I meant when I say I found a surprise in my family history.
Working With the Data You Get Back
One thing that trips people up is the triangulation process. Just because two people share DNA with you doesn't mean they share it with each other on the same segment. Without triangulating against a third match, you can't confirm whether the shared DNA comes from your mother's side or your father's side. I spent probably three weekends building a proper triangulation group for one segment that ended up revealing a great-uncle I'd never known about because my grandmother had kept that side of the family entirely out of conversation. Here's the workflow I settled on. Upload your raw DNA file to GEDmatch and FamilyTreeDNA if you haven't already. Use their chromosome browsers to pull your top matches. For matches above 20 centimorgans, check whether they also match your known relatives. If they don't, note that. It tells you which side of the family you're looking at. Then cross-reference with public genealogy trees on Ancestry. Not all matches will have trees, but the ones that do can quickly confirm or rule out a connection. The free tier on AncestryDNA gives you a fairly limited list of matches before it pushes you toward a paid subscription. You can still do meaningful work without paying, but you'll hit a wall around the forty-match mark on the free view. Uploading to MyHeritage or 23andMe gives you a larger pool of potential matches at no extra cost. I uploaded to three separate platforms and ended up with roughly double the matches I'd have had from a single test.
Things That Go Wrong and How to Fix Them
Let me tell you about a problem I ran into that probably won't be obvious unless you've actually hit it. I had a match showing up as a close relative — the algorithm was calling it a first-cousin relationship based on shared centimorgans. But when I looked at the chromosome data, the shared segment was tiny and only appeared on one chromosome arm. That's a false positive signal. It happens more often than you'd expect with certain endogamous populations or when the reference panels the company uses are thin in your ancestry group. The workaround is to look at the segment length and the number of matching SNPs, not just the total centimorgan value. If a supposed close relationship is backed by only one or two short segments below 7 centimorgans, treat it as unreliable until you find additional corroborating matches. I've seen this confuse people into building entire family trees around a single erroneous match. It took me about four hours to realize the first-cousin designation was a red herring by pulling the segment details and checking the match list against my other confirmed relatives. Another issue is phasing. Most of the big testing companies don't phase your DNA by default, which means you can't easily tell which side a shared segment comes from. You can work around this by uploading to a service that supports phasing, or by using your parents' DNA if you have access to it. Having a parent tested essentially solves the side-of-family problem overnight, and it only costs the same as any other kit. I wish I'd done that before spending a year trying to figure things out manually.
Get the Full Details

What This Approach Actually Misses
DNA testing doesn't solve everything. If your ancestors migrated before the mid-1800s, you're going to hit a brick wall no matter how many kits you order. Church records, census data, and civil registration documents are still necessary. The DNA piece tells you who your genetic relatives are. It doesn't automatically give you names, dates, or places. Those come from the traditional paper trail, and you have to build it yourself. There's also the matter of privacy. Your DNA data is permanent. Once you've uploaded it, you can't retract it from every database that processes it. Companies change their terms of service periodically. I found that out the hard way when one of the platforms I'd uploaded to announced they were sharing anonymized data with pharmaceutical partners. I deleted my account and all associated data at that point, but I had already been using the information for months. It's something to factor in before you commit to any single service. And the ethnicity estimates themselves are rough approximations. They're useful as a starting point, not as a definitive answer. I've watched people treat a 2% result in a particular region as gospel when the confidence interval for small percentages is enormous. The companies update these estimates regularly as their reference populations grow. The numbers you see today will likely look different in a year or two. Don't build your narrative around a single ethnic breakdown. Use it as a directional hint and move on to the matches, which are more stable and more useful.
If you're just getting started, the free tools on GEDmatch and the basic chromosome browser on FamilyTreeDNA are enough to begin sorting through your results. You don't need paid subscriptions or expensive consulting services unless you're dealing with especially complex cases. The biggest mistake I see people make is stopping at the ethnicity page and never actually looking at the raw match data. That's where the real information lives. The surprise in my family history came from reading the segments, not the summary report.