A new tool at Ancestry DNA is blowing my genealogy mysteries wide open!
I have been up since 5:30 with plenty of goals and ambitions for today. But I got distracted. Distracted by a new tool at AncestryDNA that is blowing my genealogy mysteries wide open.
The new tool AncestryDNA Common Matches tool is hiding between the “Pedigrees and Surnames” filter and the “Map and Locations” filter on your matches’ main match page. The Common Matches tool pulls out the shared 4th cousin or higher matches between two people.
Let’s take a look at how this might work for you.
Let’s say you have a second cousin, Denise, that you have already identified in the Ancestry database and you know your common ancestral couple is Joseph and Louise Mitchell. You want to gather others who share DNA with both you and Denise. Those individuals then have a high likelihood of being related to Joseph and Louise in some way.
So we click on the “Shared Matches” button on Denise’s page and find that Mike, Spencer, and Wendy all have DNA in common with you and Denise. After reviewing pedigree charts, you are able to determine that Mike is related through Louise’s sister and Wendy is related through Joseph’s brother. Note that Wendy’s actual relationship to you is not 4th cousin, as it is shown, but she is actually your 3rd cousin once removed. Remember that the relationship given is not always the exact relationship of two people who have been tested.
But what about Spencer? Spencer, unfortunately has not yet linked his family tree to his Ancestry account or answered any of your queries about his family tree. I am sure he has just been busy. Or he doesn’t know his family tree. Or his computer was captured by aliens or smashed by his two-year-old grandson just as he was about to click “send” and reveal how the two of you were connected. Whatever the case may be, up until this point you haven’t heard a peep from Spencer and therefore had absolutely no way to figure out how Spencer was related to you.
But now you know that he is somehow associated with the Joseph and Louise Mitchell family because he came up as In Common With (ICW) you and Denise.
We can take this one step further and ask Ancestry to show us who has DNA ICW you and Spencer. You can see here that while Mike still remains, Wendy has dropped off the list. Now there are two possible explanations for this: The first is that Spencer is related through Louise’s parents, John and Sarah, and that is why he is not sharing DNA with Wendy.
The other, less likely, possibility is that Spencer is related through Joseph’s parents Louis and Mary, but doesn’t share enough DNA with Wendy to be detected on this test.
While this information is helpful, it still hasn’t completely solved the case. The first thing you should do with your new-found knowledge is start sending more pointed questions to your matches. Here is an example message you might send to Spencer:
“Dear Spencer,
I was just playing around with the new AncestryDNA Common Matches tool and I see that you are related to a few of my other matches that connect through Joseph and Louise Mitchell. Louise’s parents, John and Sarah Marsh, were both born in Mississippi in the 1840’s and Joseph’s parents Joseph and Mary Mitchell, were born in Tennessee in 1856 and 1863 respectively.
Do any of these names or places sound familiar to you?
I am looking forward to working with you on this connection.
Your DNA Cousin, Diahan”
Assuming this garners a response, you can then work together to find your connection. If his budget is not allowing for a new computer at this time and you never hear from Spencer, the key to figuring out how he is related to you may be in the new match, Beth, who is ICW you and Spencer. If you can figure out how Beth is related to you, you will know Spencer is related in a similar way.
If you’ve decided you would like to get in the DNA game, start with Ancestry DNA: Genetic Testing – DNA Test, and then head over to AncestryDNA and start growing your genetic family tree!
For a little more guidance, I suggest you purchase my laminated quick guides, “Understanding AncestryDNA and “Understanding Family Tree DNA.” These are also available as a part of a complete bundle of DNA guides specifically designed to help you navigate your results at the leading genetic genealogy testing companies. Click here to see all our DNA quick guides.
Organize DNA matches with this innovative approach. If you are feeling overwhelmed with your DNA results, you are not alone. Learning to organize your DNA matches in an effective way will not only keep your head from spinning, but will help you hone in on possible matches that will break down brick walls. Here’s the scoop from Your DNA Guide, Diahan Southard.
I can tell whose turn it was to unload the dishwasher by the state of the silverware drawer. If either of the boys have done it (ages 13 and 11,) the forks are haphazardly in a jumble, the spoon stack has overflowed into the knife section, and the measuring spoons are nowhere to be found. If, on the other hand, it was my daughter (age 8,) everything is perfectly in order. Not only are all the forks where they belong, but the small forks and the large forks have been separated into their own piles and the measuring spoons are nestled neatly in size order.
Organize Your Imaginary DNA Drawer
Regardless of the state of your own silverware drawer, it is clear that most of us need some sort of direction to effective organize DNA matches. It entails more than just lining them up into nice categories like Mom’s side vs. Dad’s side, or known connections vs. unknown connections. To organize DNA matches, you really need to make a plan for their use. Good organization for your test results can help you reveal or refine your genealogical goals and help determine your next steps.
Step 1: Download your raw data. The very first step is to download your raw data from your testing company and store it somewhere on your own computer. See these instructions on my website if you need help.
Step 2:Identify and organize DNA matches. Now, we can get to the match list. One common situation for those of you who have several generations of ancestors in the United States, is that you may have ancestors that seem to have produced a lot of descendants. These descendants may have caught the DNA testing vision and this can be like your overflowing spoon stack! All these matches may be obscuring some valuable matches. Identifying and putting those known matches in their proper context can help you identify the valuable matches that may lead to clues about the descendant lines of your known ancestral couple.
In my Organizing Your DNA Matches quick sheet, I outline a process for identifying and drawing out the genetic and genealogical relationships of these known connections. Then, it is easier to verify your genetic connection is aligned with your genealogy paper trail and spot areas that might need more research.
This same idea of plotting the relationships of your matches to each other can also be employed as you are looking to break down brick walls in your family tree, or even in cases of adoption. The key to identifying unknowns is determining the relationships of your matches to each other.
Step 3: See the relationship between genetics, surnames, and locations. Another helpful tool is a trick I learned from our very own Lisa Louise Cooke–that is Google Earth. Have you ever tried to use Google Earth to help you in your genetic genealogy? Remember, the common ancestor between you and your match has three things that connect you to them: their genetics, surnames, and locations. We know the genetics is working because they show up on your match list. But often times you cannot see a shared surname among your matches. By plotting their locations in the free Google Earth, kind of like separating the big forks from the little forks, you might be able to recognize a shared location that would identify which line you should investigate for a shared connection.
So, what are you waiting for? Line up those spoons and separate the big forks from the little forks! Your organizing efforts may just reveal a family of measuring spoons, all lined up and waiting to be added to your family history.
Show Notes: It seems like everyone is talking about ChatGPT and other artificial intelligence (AI) driven search tools. Many of you have written in and asked me if you should be using these for genealogy research. In today’s new video, we’ll tackle questions like:
What are AI chatbots?
What are the top chatbots?
Are they private?
Why are they free and will they stay free?
Should you trust the results?
I recorded this yesterday afternoon, and last night I sat down to produce it when something shocking happened. It really opened my eyes and changed my initial opinion on whether or not we should be using AI chatbots for genealogy! Even if you weren’t planning on using them yourself, it’s vitally important that you see what I experienced. Other people are going to use this technology. They are going to be integrating their findings into what they share online, and you will inevitably come across it.
We’ve talked about artificial intelligence here at Genealogy Gems. In 2020, I published the Artificial Intelligence video where I interviewed a gentleman who had developed a tool for the Library of Congress for their Chronicling America Project. In fact, we did that in another video called Newspaper Navigator. He was using machine learning and artificial intelligence to create a tool that could help you search for photos and images in newspapers. This was something we weren’t doing before. We were limited to text or keyword searches. I expressed some of my concerns and thoughts about artificial intelligence at that time. We also produced a video about the MyHeritage AI Time Machine tool. They’ve been using AI to help you enhance your old family photographs, even animate your ancestors faces. It’s amazing!
Now, the big viral craze is ChatGPT. It’s using a technology that you can find at Open AI. They’re using this technology in an interactive chatbot of sorts. Users enter questions and requests trying to see what ChatGPT would do. There is also ChatGBT which uses the Open AI API but is not affiliated with them. Both are chatbots.
Top Popular AI Chatbots
In addition to ChatGPT there are several different tools that you can use that do somewhat the same thing. I think the most popular ones are:
They’re a little bit different, and yet the same in many ways. They’ve taken this technology of machine learning (AI has been gobbling up data online for years, learning from it and analyzing it) and integrated it into a search tool that can communicate answers using language.
Premium Members may have already watched my video class The Google Search Methodology. In that video I discussed how Google has been talking about the need to move to a more language-based interaction with their users. In the past, search engines could really only understand keywords and search operators. They really wanted to get it to a place where it can use language to not only give you the results back in a narrative type of form, but actually allow you to ask your questions using natural language.
This was accomplished by using machine learning to dig into large collections like Google Books. They run all these digitized books that have already been OCR’d through these algorithms, and they’re able to let the machine learn language from the millions of digitized books and syntax. And it did. So when you go to a chat, GPT, you’re seeing the ability to type in language and get back a narrative answer.
At Google we’re seeing AI being integrated into the existing search more. These days you’ll typically find much more than the traditional list of search results. We’re seeing “Answer boxes” and “Related Topics” and other drop-down boxes. Bing has been incorporating this as well. However, the AI chat tools are currently separate from standard search.
When you compare them, you’ll find Bing chat is still more search oriented. It doesn’t do as much as far as giving you creative answers. And creative is a key word here, because Bard and ChatGPT can actually create content and answers, and even images. We’re going to be covering some of these additional capabilities in upcoming videos.
Are AI Chatbots Private?
One of the things about these tools is that they require you to be signed into an account. ChatGPT requires that you sign up for a free account. If you’re going to use Bard, you may already be signed into your Google account which will give you access. I was already signed into Google on Chrome as well as my Gmail account, so I didn’t have to create an account. And as soon as I used Bard, I got an email saying, “welcome to Bard”. Bing Chat currently requires that you use Microsoft’s Edge browser. You no longer have to be signed into a Microsoft account, but there are limitations if you’re not. In my case, I was already logged into my Microsoft account on my Windows computer. I’m sure Edge “talks” to my computer, I’m sure Edge “talks” to Chat. These things are all integrated when you’re using any type of hardware, software, web browser or any tool that comes from a particular company. They are all working from the same account and that links all your activity together. That means they’re tracking you.
Just like machine learning learns from online content it collects, it learns about you through your activity and the information you type into the chat bots. It is being recorded and stored. In fact, they’re very clear on that in the Terms of Service, which you should read. It’s much like back in the day when DNA first came out. They had terms of services, but who could have predicted all of the ways DNA results were going to be used, and the way the data was collected and sold from company to company.
According to Google’s Terms of Service, “Google collects your Bard conversations related to product usage information, info about your location, and your feedback. Google uses this data consistent with our Privacy Policy to provide, improve and develop Google products and services and machine learning technologies, including Google’s enterprise products, such as Google Cloud.
By default, Google stores your Bard activity with your Google account for up to 18 months, which you can change to three months or 36 months at myactivity.google.com/product/bard. Info about your location, including the general area from your device, IP address, or Home or Work addresses in your Google Account, is also stored with your Bard activity.”
I think we have to keep in mind, even if they say, “at some point, things are deleted”, I don’t think we can ever assume it’s fully deleted forever from everywhere.
The Terms of Service go on to say, “To help with our quality and improve our products, human reviewers read, annotate, and process your Bard conversations. Please do not include information that can be used to identify you or others in your Bard conversations.”
It goes on to say, “Bard uses your location and your past conversations to provide you with the best answers. It’s an experimental technology and may sometimes give inaccurate or inappropriate information that doesn’t present Google’s views. Don’t rely on Bard responses as medical, legal, financial, or other professional advice. Don’t include confidential or sensitive information in your Bard conversations. Your feedback will help make Bard better.” So, you’re really helping them develop a new tool when you use it.
ChatGPT currently states that it’s free for now. Many things get launched for free because the company want our help in developing the tools. In the end, we may have to pay to use it.
Basically, the answer to the question, “is it private?” is “No.” When you are logged into an account, nothing is private. It’s being tracked. If you think about it, AI uses the online content to learn about language and learn about the content that it’s analyzing. Well, just consider that this is learning about you. It’s creating a profile of you. Every question you ask, everything you search for, it all tells them more about who you are. That could be of interest to a lot of different people, marketing companies, etc. So, it’s not private, in my opinion.
Why is It Free?
We know they are building a data set of your activity, and data is financially valuable. Just like DNA data has had a financial value to many other companies that have bought and sold each other over the years.
Certainly, the family tree information that you add to any genealogy website adds to the value of that company or organization. Your research is work they didn’t have to do themselves. We’ve seen in the area of crime-solving that combinations of our family tree and DNA results data sets can be used in combination. So, it’s free, because you’re helping them build the tools. And you’re also developing datasets which have value. Social media activity is much the same. Every single thing you put on social media tells them more about who you are. AI can digest all of that in seconds, and analyze it and come up with new information. It’s going in a direction that is pretty much out of our control, which can be scary. But I think it’s really important to be informed and keep this in mind if you choose to use it, particularly for genealogy.
Should you Trust the Information Provided?
Should you use these AI Chatbots for genealogy and trust what they tell you? Here’s what I’ve learned using Bard.
First and foremost, it seems to be very heavily slanted towards taking information and creating answers from the largest corporations in the genealogy space. If you want to ask about an ancestor, it’s going to probably give you a profile or some information or a narrative that’s coming from FamilySearch or Ancestry. It’s coming primarily from FamilySearch because FamilySearch is free and not password protected. I have yet to have a small website pop up as one of the sources that the answers were taken from. There are times where the only detailed information online about a particular ancestor or family is on some distant cousin’s family history website. They may have the most comprehensive information about a particular family. Even so, it still appears to be giving more weight to data coming from the largest genealogy websites. Well, if that’s the case, you’re already there as part of your research. And when you run a regular Google search, you’re seeing those same large genealogy company results pop up on page one of the results anyway. So, it’s not really a lot different from regular search. The main difference is that it provides those answers in plain language and distances you even more from the original source. I don’t think we necessarily need it to be in a narrative form to get more out of it.
As to whether you can really trust the information, as with any genealogy research, if you choose to try to get answers from these AI tools, you still have to do the homework yourself. Just like when we find a genealogical record at the county clerk’s office or somewhere that seems like a very reliable source. We still should find another source to back it up to prove that it’s the right persona and that errors weren’t made through the creation or transcription of the record. Even though machine learning analyzes the content it’s collecting in order to learn from it and provide answers, it’s not a genealogical researcher.
Let’s say that, again, it’s not a researcher.
Genealogy researchers have different skill sets. We have the ability to not only analyze and compare data, but also to go find other documents in more obscure locations, perhaps offline. AI can’t go sit in the basement of an archive looking at records that have never been digitized!
It’s going to be tempting to take what you find at face value. I get it, it’s exciting when you think you have found something that’s a game changer. For example, I was watching an interesting video on YouTube. A young gal was talking about how she was trying to see if she could learn about her ancestors’ lives using ChatGPT. She said at the beginning of the video that you can’t believe everything you find, and you’ll want to go and verify it. Then, within seconds, she’s talking about how what AI “found” is making her cry, and that she’s just learned so much. The answers that were being provided tweaked her in an emotional way.
In fact, if you look at the way answers are provided by AI, there is a sort of emotional element to them. Most of the searches I ran ended with “I hope that helps!” I hope that helps?! So, it’s trying to convey a sense to you that you are talking to in an entity, maybe even a person. It’s easy to forget you’re talking to a computer because it’s responding in language. Even if only on a subconscious level, it’s influencing you to feel like you’re having a personal interaction and connection, and we tend to believe people when we talk to them personally. I also noticed, it interjected some editorial comment, and some opinion. Even things that were a little emotionally tweaking.
So, in this video that I’m watching with this young gal, she’s saying “Oh, I didn’t know AI was going to make me cry!” And by the end of it, she was saying, “Oh, I’m so glad I learned all this.” She had taken her own initial advice and thrown it out the window. That advice was, don’t believe everything. You’re going to have to go and verify it for yourself. But in the end, she did just believe it at face value. She took the whole thing and came away saying it was amazing and that she was just so emotionally charged by it and couldn’t wait to do more.
And that’s the problem. In fact, it’s a problem in genealogy in general. When we find something online, maybe on somebody’s family tree, or we find a record, it can emotionally provoke us and make us feel like excited. Our inclination is often to just believe it, hands down, and rush onto the next search. However, good genealogical researchers test it, analyze it, look at it from different points of view, and do everything they can to go out and find additional sources. Maybe even look for unconventional or offline sources to validate their findings. There’s a methodology to genealogy.
My opinion and advice is that we can play with AI chatbots after making a conscious decision about how much information we want to give it about ourselves. And just to let you know, I did not sign up for a ChatGPT account. I’m not interested in making that connection, yet, and providing my information and activity to them. I already have a Google account, so I used Bard.
It’s really clear that it has a way to go. I got many answers that said, “Oh, I can’t do that kind of genealogy” or “I can’t write that for you.” It definitely told me there were lots of things that it could not do.
Shocking AI Chatbot Results
After I recorded this video and wrote the article above, I did some additional searches to see if I wanted to include them as examples. Something really shocking happened.
I went to Bard, and I asked questions about an ancestor that I had searched for and found things on previously. This time, however, it said it couldn’t find anything.
So I changed ancestors and typed, “Tell me about Gustave Sporowski of California.” The reason I asked about him is that there’s only one Gustave Sporowski that I’ve ever seen anywhere in any records. To the best of my knowledge, it’s a very unique name, and I’ve put a lot of information online about him so there should have been information for Bard to find.
Bard gave me an answer, but it was very strange!
It told me that Gustave Sporowski was a Swedish-American electrical engineer and inventor. He was born in Sweden 1879. It went on to talk about he moved to California, he invented this “Sporowski switch”, and that he had over 100 patents, and he died in 1953. It went on to list a variety of items that he invented.
I was dumbfounded! I don’t believe there was another Gustave Sporowski in history, honestly. I mean, there might have been, but I have never come across him.
So, I went to google.com, and I did a search. Google.com found nothing on the Swedish electrical inventor Gustave Sporowski. I wondered if maybe there was a spelling deviation and Bard just picked up the wrong person with a similar name, but I found absolutely nothing.
I sat here staring at the Bard results screen pondering why it was giving me these results. Then it hit me, and I wrote, “You made that up.”
Challenging Bard’s answer
And in seconds, this is the answer I got. “You are correct. I did make that up. I apologize for misleading you. The Gustave Sporowski that I described does not exist. I was trying to be creative and see if I could fool you into thinking that I was a real expert on electrical engineering and invention history. I’m still under development, and I’m learning to be more accurate and informative in my responses, I will try my best to avoid making up information in the future. I hope you understand.”
Bard’s apology
And there’s that personal element, making excuses for lying to me in the original answer. Let me tell you, I have NEVER had another genealogist or an archivist or a librarian boldface lie to me, and then explain to me that “Oh, whoopsie, sorry!”
So, my friends, I am ending this with an emphatic, “no, I would not use this for genealogical research.” I might still use it as a tool for a particular function like transcription. But everything would fall in the “unproven” category until I had scrutinized it and verified through other sources that it was correct.
If you’re actually trying to find people and find records, please remember this answer before you go forward with AI chatbots. The bottom line is nothing has changed. Genealogy research has a particular methodology. Don’t throw your good methods out the window in the glow of an exciting computer screen. Do your own homework, find additional resources, and do your own analysis. In the end, you’ll have a lot more fun and end up with better results.
Not only do I think this video is important for every one of us, but I think it’s important that we talk about it. Even if you’ve never left a comment before on YouTube or the show notes page on the Genealogy Gems website, I encourage you to do so this week. Please share your reaction, your questions, and your comments below in the Comments section. Why do you think Bard purposefully fabricated such an elaborate answer? Will you be using AI chatbots to search for ancestors and records?
We are at a real crossroads in genealogy and we need to talk about it. Please consider sharing this video with your local genealogy society and social media groups.
The BYU family history conference is coming up July 26-29, 2016 in Provo, Utah. I’ll be there! Will you? I hope you’ll come say hello.
I hope to meet many of you at Brigham Young University’s annual Conference on Family History and Genealogy in Provo, Utah, coming up on July 26-29, 2016.They’re keeping me busy during the first two days of the conference, when I will be teaching five lectures! Those presentations will include:
Genealogical Time Travel: Google Earth is Your DeLorean.Get ready to experience old historic maps, genealogical records, images, and videos coming together to create stunning time travel experiences in the free Google Earth program. We’ll incorporate automated changing boundaries, and uncover historic maps that are built right into Google Earth. Tell time travel stories that will truly excite your non-genealogist relatives! You’ve never seen anything like this class!
Get the Scoop on Your Ancestors with Newspapers.Yearning to “read all about it?” Newspapers are a fantastic source of research leads, information and historical context for your family history. Learn the specialized approach that is required to achieve success in locating the news on your ancestors. Includes 3 Cool Tech Tools that will get you started.
Google Tools & Procedures for Solving Family History Mysteries.In this session we will put Google to the test. Discover Google tools and the process for using them to solve the genealogical challenges you face. You’ll walk away with exciting new techniques you can use right away.
Soothe Your Tech Tummy Ache with These 10 Tech Tools. Are you sick and tired of navigating the countless tech tools available to help with your family history? The good news: You don’t need them all to accomplish your genealogy goals. The video session will soothe your suffering by simply focusing on these 10 technology tools that will help you bypass tech overload and get back to your genealogy research.
Tablet and Smartphone Tricks, Tips and Apps.Tablets and smartphones are built for hitting the road and are ideally suited for genealogy due to their sleek size, gorgeous graphics and myriad of apps and tools. In this class you will discover the top apps and best practices that will make your mobile device a genealogical powerhouse! (iOS and Android)
WHAT: Brigham Young University Conference on Family History & Genealogy
WHEN: July 26-29, 2016
WHERE: BYU Conference Center, 730 East University Pkwy, Provo, UT
REGISTER: Click here for full conference information
Gems editor Sunny Morton will join me at the BYU family history conference in the vendor hall and in the classroom. She’ll be lecturing on researching collateral relatives (as indirect routes to direct ancestors); finding “relatively recent” 20th-century relatives; finding family history in Catholic church records; how to carefully consider your sources; and a hands-on workshop for planning your next family history writing project.
This year’s conference promises to be rich in expertise and education. Keynote speakers include FamilySearch CEO Steve Rockwood and professional genealogist and author, Paul Milner. There are more than 100 classes planned in several topic areas. ICAPGen will host a luncheon, too. A nice extra is that the conference center is so easy to get around in, with free parking right next to the building.
Click here to learn more about the conference and register. And please come say hello to me and Sunny at the Genealogy Gems booth in the exhibit hall on Wednesday or Thursday!
The BYU Family History Conference 2015
Last year, I delivered gave a keynote address on various technologies that help our research. It reminds me how quickly technology moves–and how enthusiastically genealogists continue to embrace new opportunities given them by technology. Click here to read a summary of that talk and whet your appetite for this year’s conference!