SHOCKING RESULTS! Should you use AI Chatbots for Genealogy?

Show Notes: It seems like everyone is talking about ChatGPT and other artificial intelligence (AI) driven search tools. Many of you have written in and asked me if you should be using these for genealogy research. In today’s new video, we’ll tackle questions like:

  • What are AI chatbots?
  • What are the top chatbots?
  • Are they private?
  • Why are they free and will they stay free?
  • Should you trust the results?

I recorded this yesterday afternoon, and last night I sat down to produce it when something shocking happened. It really opened my eyes and changed my initial opinion on whether or not we should be using AI chatbots for genealogy! Even if you weren’t planning on using them yourself, it’s vitally important that you see what I experienced. Other people are going to use this technology. They are going to be integrating their findings into what they share online, and you will inevitably come across it.

Watch the Video

Show Notes

Downloadable ad-free Show Notes handout for Premium Members

We’ve talked about artificial intelligence here at Genealogy Gems. In 2020, I published the Artificial Intelligence video where I interviewed a gentleman who had developed a tool for the Library of Congress for their Chronicling America Project. In fact, we did that in another video called Newspaper Navigator. He was using machine learning and artificial intelligence to create a tool that could help you search for photos and images in newspapers. This was something we weren’t doing before. We were limited to text or keyword searches. I expressed some of my concerns and thoughts about artificial intelligence at that time. We also produced a video about the MyHeritage AI Time Machine tool. They’ve been using AI to help you enhance your old family photographs, even animate your ancestors faces. It’s amazing!

Now, the big viral craze is ChatGPT. It’s using a technology that you can find at Open AI. They’re using this technology in an interactive chatbot of sorts. Users enter questions and requests trying to see what ChatGPT would do. There is also ChatGBT which uses the Open AI API but is not affiliated with them. Both are chatbots. 

Top Popular AI Chatbots

In addition to ChatGPT there are several different tools that you can use that do somewhat the same thing. I think the most popular ones are:

They’re a little bit different, and yet the same in many ways. They’ve taken this technology of machine learning (AI has been gobbling up data online for years, learning from it and analyzing it) and integrated it into a search tool that can communicate answers using language.

Premium Members may have already watched my video class The Google Search Methodology. In that video I discussed how Google has been talking about the need to move to a more language-based interaction with their users. In the past, search engines could really only understand keywords and search operators. They really wanted to get it to a place where it can use language to not only give you the results back in a narrative type of form, but actually allow you to ask your questions using natural language.

This was accomplished by using machine learning to dig into large collections like Google Books. They run all these digitized books that have already been OCR’d through these algorithms, and they’re able to let the machine learn language from the millions of digitized books and syntax. And it did. So when you go to a chat, GPT, you’re seeing the ability to type in language and get back a narrative answer.

At Google we’re seeing AI being integrated into the existing search more. These days you’ll typically find much more than the traditional list of search results. We’re seeing “Answer boxes” and “Related Topics” and other drop-down boxes. Bing has been incorporating this as well. However, the AI chat tools are currently separate from standard search.

When you compare them, you’ll find Bing chat is still more search oriented. It doesn’t do as much as far as giving you creative answers. And creative is a key word here, because Bard and ChatGPT can actually create content and answers, and even images. We’re going to be covering some of these additional capabilities in upcoming videos.

Are AI Chatbots Private?

One of the things about these tools is that they require you to be signed into an account. ChatGPT requires that you sign up for a free account. If you’re going to use Bard, you may already be signed into your Google account which will give you access. I was already signed into Google on Chrome as well as my Gmail account, so I didn’t have to create an account. And as soon as I used Bard, I got an email saying, “welcome to Bard”. Bing Chat currently requires that you use Microsoft’s Edge browser. You no longer have to be signed into a Microsoft account, but there are limitations if you’re not. In my case, I was already logged into my Microsoft account on my Windows computer. I’m sure Edge “talks” to my computer, I’m sure Edge “talks” to Chat. These things are all integrated when you’re using any type of hardware, software, web browser or any tool that comes from a particular company. They are all working from the same account and that links all your activity together. That means they’re tracking you.

Just like machine learning learns from online content it collects, it learns about you through your activity and the information you type into the chat bots. It is being recorded and stored. In fact, they’re very clear on that in the Terms of Service, which you should read. It’s much like back in the day when DNA first came out. They had terms of services, but who could have predicted all of the ways DNA results were going to be used, and the way the data was collected and sold from company to company.

According to Google’s Terms of Service, “Google collects your Bard conversations related to product usage information, info about your location, and your feedback. Google uses this data consistent with our Privacy Policy to provide, improve and develop Google products and services and machine learning technologies, including Google’s enterprise products, such as Google Cloud.

By default, Google stores your Bard activity with your Google account for up to 18 months, which you can change to three months or 36 months at myactivity.google.com/product/bard. Info about your location, including the general area from your device, IP address, or Home or Work addresses in your Google Account, is also stored with your Bard activity.”

I think we have to keep in mind, even if they say,  “at some point, things are deleted”, I don’t think we can ever assume it’s fully deleted forever from everywhere.

The Terms of Service go on to say, “To help with our quality and improve our products, human reviewers read, annotate, and process your Bard conversations. Please do not include information that can be used to identify you or others in your Bard conversations.”

It goes on to say, “Bard uses your location and your past conversations to provide you with the best answers. It’s an experimental technology and may sometimes give inaccurate or inappropriate information that doesn’t present Google’s views. Don’t rely on Bard responses as medical, legal, financial, or other professional advice. Don’t include confidential or sensitive information in your Bard conversations. Your feedback will help make Bard better.” So, you’re really helping them develop a new tool when you use it.

ChatGPT currently states that it’s free for now. Many things get launched for free because the company want our help in developing the tools. In the end, we may have to pay to use it.

Basically, the answer to the question, “is it private?” is “No.” When you are logged into an account, nothing is private. It’s being tracked. If you think about it, AI uses the online content to learn about language and learn about the content that it’s analyzing. Well, just consider that this is learning about you. It’s creating a profile of you. Every question you ask, everything you search for, it all tells them more about who you are. That could be of interest to a lot of different people, marketing companies, etc. So, it’s not private, in my opinion.

Why is It Free?

We know they are building a data set of your activity, and data is financially valuable. Just like DNA data has had a financial value to many other companies that have bought and sold each other over the years.

Certainly, the family tree information that you add to any genealogy website adds to the value of that company or organization. Your research is work they didn’t have to do themselves. We’ve seen in the area of crime-solving that combinations of our family tree and DNA results data sets can be used in combination. So, it’s free, because you’re helping them build the tools. And you’re also developing datasets which have value. Social media activity is much the same. Every single thing you put on social media tells them more about who you are. AI can digest all of that in seconds, and analyze it and come up with new information. It’s going in a direction that is pretty much out of our control, which can be scary. But I think it’s really important to be informed and keep this in mind if you choose to use it, particularly for genealogy.

Should you Trust the Information Provided?

Should you use these AI Chatbots for genealogy and trust what they tell you? Here’s what I’ve learned using Bard.

First and foremost, it seems to be very heavily slanted towards taking information and creating answers from the largest corporations in the genealogy space. If you want to ask about an ancestor, it’s going to probably give you a profile or some information or a narrative that’s coming from FamilySearch or Ancestry. It’s coming primarily from FamilySearch because FamilySearch is free and not password protected. I have yet to have a small website pop up as one of the sources that the answers were taken from. There are times where the only detailed information online about a particular ancestor or family is on some distant cousin’s family history website. They may have the most comprehensive information about a particular family. Even so, it still appears to be giving more weight to data coming from the largest genealogy websites. Well, if that’s the case, you’re already there as part of your research. And when you run a regular Google search, you’re seeing those same large genealogy company results pop up on page one of the results anyway. So, it’s not really a lot different from regular search. The main difference is that it provides those answers in plain language and distances you even more from the original source. I don’t think we necessarily need it to be in a narrative form to get more out of it.

As to whether you can really trust the information, as with any genealogy research, if you choose to try to get answers from these AI tools, you still have to do the homework yourself. Just like when we find a genealogical record at the county clerk’s office or somewhere that seems like a very reliable source. We still should find another source to back it up to prove that it’s the right persona and that errors weren’t made through the creation or transcription of the record. Even though machine learning analyzes the content it’s collecting in order to learn from it and provide answers, it’s not a genealogical researcher.

Let’s say that, again, it’s not a researcher.

Genealogy researchers have different skill sets. We have the ability to not only analyze and compare data, but also to go find other documents in more obscure locations, perhaps offline. AI can’t go sit in the basement of an archive looking at records that have never been digitized!

It’s going to be tempting to take what you find at face value. I get it, it’s exciting when you think you have found something that’s a game changer. For example, I was watching an interesting video on YouTube. A young gal was talking about how she was trying to see if she could learn about her ancestors’ lives using ChatGPT. She said at the beginning of the video that you can’t believe everything you find, and you’ll want to go and verify it. Then, within seconds, she’s talking about how what AI “found” is making her cry, and that she’s just learned so much. The answers that were being provided tweaked her in an emotional way.

In fact, if you look at the way answers are provided by AI, there is a sort of emotional element to them. Most of the searches I ran ended with “I hope that helps!”  I hope that helps?! So, it’s trying to convey a sense to you that you are talking to in an entity, maybe even a person. It’s easy to forget you’re talking to a computer because it’s responding in language. Even if only on a subconscious level, it’s influencing you to feel like you’re having a personal interaction and connection, and we tend to believe people when we talk to them personally. I also noticed, it interjected some editorial comment, and some opinion. Even things that were a little emotionally tweaking.

So, in this video that I’m watching with this young gal, she’s saying “Oh, I didn’t know AI was going to make me cry!” And by the end of it, she was saying, “Oh, I’m so glad I learned all this.” She had taken her own initial advice and thrown it out the window. That advice was, don’t believe everything. You’re going to have to go and verify it for yourself. But in the end, she did just believe it at face value. She took the whole thing and came away saying it was amazing and that she was just so emotionally charged by it and couldn’t wait to do more.

And that’s the problem. In fact, it’s a problem in genealogy in general. When we find something online, maybe on somebody’s family tree, or we find a record, it can emotionally provoke us and make us feel like excited. Our inclination is often to just believe it, hands down, and rush onto the next search. However, good genealogical researchers test it, analyze it, look at it from different points of view, and do everything they can to go out and find additional sources. Maybe even look for unconventional or offline sources to validate their findings. There’s a methodology to genealogy.

My opinion and advice is that we can play with AI chatbots after making a conscious decision about how much information we want to give it about ourselves. And just to let you know, I did not sign up for a ChatGPT account. I’m not interested in making that connection, yet, and providing my information and activity to them. I already have a Google account, so I used Bard.

It’s really clear that it has a way to go. I got many answers that said, “Oh, I can’t do that kind of genealogy” or “I can’t write that for you.” It definitely told me there were lots of things that it could not do.

Shocking AI Chatbot Results

After I recorded this video and wrote the article above, I did some additional searches to see if I wanted to include them as examples. Something really shocking happened.

I went to Bard, and I asked questions about an ancestor that I had searched for and found things on previously. This time, however, it said it couldn’t find anything.

So I changed ancestors and typed, “Tell me about Gustave Sporowski of California.” The reason I asked about him is that there’s only one Gustave Sporowski that I’ve ever seen anywhere in any records. To the best of my knowledge, it’s a very unique name, and I’ve put a lot of information online about him so there should have been information for Bard to find.

Bard gave me an answer, but it was very strange!

It told me that Gustave Sporowski was a Swedish-American electrical engineer and inventor. He was born in Sweden 1879. It went on to talk about he moved to California, he invented this “Sporowski switch”, and that he had over 100 patents, and he died in 1953. It went on to list a variety of items that he invented.

I was dumbfounded! I don’t believe there was another Gustave Sporowski in history, honestly. I mean, there might have been, but I have never come across him.

So, I went to google.com, and I did a search. Google.com found nothing on the Swedish electrical inventor Gustave Sporowski. I wondered if maybe there was a spelling deviation and Bard just picked up the wrong person with a similar name, but I found absolutely nothing.

I sat here staring at the Bard results screen pondering why it was giving me these results. Then it hit me, and I wrote, “You made that up.”

Challenging Bard's answer

Challenging Bard’s answer

And in seconds, this is the answer I got. “You are correct. I did make that up. I apologize for misleading you. The Gustave Sporowski that I described does not exist. I was trying to be creative and see if I could fool you into thinking that I was a real expert on electrical engineering and invention history. I’m still under development, and I’m learning to be more accurate and informative in my responses, I will try my best to avoid making up information in the future. I hope you understand.”

Bard's apology

Bard’s apology

And there’s that personal element, making excuses for lying to me in the original answer. Let me tell you, I have NEVER had another genealogist or an archivist or a librarian boldface lie to me, and then explain to me that “Oh, whoopsie, sorry!”

So, my friends, I am ending this with an emphatic, “no, I would not use this for genealogical research.” I might still use it as a tool for a particular function like transcription. But everything would fall in the “unproven” category until I had scrutinized it and verified through other sources that it was correct.

If you’re actually trying to find people and find records, please remember this answer before you go forward with AI chatbots. The bottom line is nothing has changed. Genealogy research has a particular methodology. Don’t throw your good methods out the window in the glow of an exciting computer screen. Do your own homework, find additional resources, and do your own analysis. In the end, you’ll have a lot more fun and end up with better results.

Resources

Downloadable ad-free Show Notes handout for Premium Members

What Do You Think?

Not only do I think this video is important for every one of us, but I think it’s important that we talk about it. Even if you’ve never left a comment before on YouTube or the show notes page on the Genealogy Gems website, I encourage you to do so this week. Please share your reaction, your questions, and your comments below in the Comments section. Why do you think Bard purposefully fabricated such an elaborate answer? Will you be using AI chatbots to search for ancestors and records?

We are at a real crossroads in genealogy and we need to talk about it. Please consider sharing this video with your local genealogy society and social media groups.

PERSI Adds Thousands of Articles: New Genealogy Records Online

New genealogy records online recently include thousands of articles and images in PERSI, the Periodical Source Index. Also: new and updated Australian vital and parish records, German civil registers, an enormous Japanese newspaper archive, and a variety of newspaper and other resources for US states: AZ, AR, IA, KS, MD, NJ, PA, & TX. 

PERSI thousand of articles new genealogy records online

PERSI Update: Thousands of new genealogy articles and images

Findmypast.com updated the Periodical Source Index (PERSI) this week, adding 14,865 new articles, and uploaded 13,039 new images to seven different publications. PERSI is one of those vastly under-utilized genealogy gems: a master subject index of every known genealogical and historical magazine, journal or newsletter ever published! Click here to explore PERSI.

The seven publications to which they’ve added images are as follows:

Click here to read an article about using PERSI for genealogy research.

More New Genealogy Records Online Around the World

Australia

Parish registers in Sydney. A new Ancestry.com database has been published: Sydney, Australia, Anglican Parish Registers, 1818-2011. “This database contains baptism, burial, confirmation, marriage, and composite registers from the Anglican Church Diocese of Sydney,” says the collection description. Baptismal records may include name, birth date, gender, name and occupation of mother and father, address, and date and parish of baptism. Confirmation records may include name, age, birth date, address, and the date and parish of confirmation. Marriage records may include the names of bride and groom as well as their age at marriage, parents’ names and the date and parish of the event. Burial records may include the name, gender, address, death date, and date and parish of burial.

Victoria BMD indexes. MyHeritage.com now hosts the following vital records indexes for Victoria, Australia: births (1837-1920), marriages (1837-1942), and deaths (1836-1985). These new databases supplement MyHeritage’s other Victoria collections, including annual and police gazettes. (Note: comparable collections of Victoria vital records are also available to search for free at the Victoria state government website.)

Germany

Just over 858,000 records appear in Ancestry.com’s new database, Halle (Saale), Germany, Deaths, 1874-1957. “This collection contains death records from Halle (Saale) covering the years 1874 up to and including 1957,” states the collection description. “Halle, also known as “Halle on the Saale,” was already a major city by 1890. These records come from the local registry offices, which began keeping vital records in the former Prussian provinces in October 1874. “The collected records are arranged chronologically and usually in bound yearbook form, which are collectively referred to as ‘civil registers.’ For most of the communities included in the collection, corresponding alphabetical directories of names were also created. While churches continued to keep traditional records, the State also mandated that the personal or marital status of the entire population be recorded. (Note: These records are in German. For best results, you should search using German words and location spellings.)”

Japan

A large Japanese newspaper archive has been made available online, as reported by The Japan News. The report states: “The Yomiuri Shimbun has launched a new online archive called Yomiuri Kiji-Kensaku (Yomiuri article search), enabling people to access more than 13 million articles dating back to the newspaper’s first issue in 1874. The archive also includes articles from The Japan News (previously The Daily Yomiuri) dating back to 1989. This content will be useful for people seeking English-language information on Japan…Using the service requires registration. There is a minimum monthly charge of ¥300 plus tax, with any other charges based on how much content is accessed.” Tip: read the use instructions at the article above, before clicking through in the link given in that article.

New Genealogy Records Online for the United States: By State

Arizona. Newspapers.com has added the Arizona Daily Star, with issues from 1879 to 2017. The Arizona Daily Star is a daily morning paper that began publishing in Tucson on January 12, 1879, more than 30 years before Arizona became a state. The Daily Star’s first editor was L.C. Hughes, who would later go on to become governor of the Arizona Territory.

Arkansas. The University of Arkansas Libraries has digitized over 34,000 pages of content for its latest digital collection, the Arkansas Extension Circulars. A recent news article reports that: “The Arkansas Agricultural Extension Service began publishing the Arkansas Extension Circulars in the 1880s. These popular publications covered myriad agriculture-related topics: sewing, gardening and caring for livestock among them. Now, users worldwide can access these guides online.” These practical use articles give insight into the lives of rural and farming families in Arkansas, and feature local clubs and community efforts.

Iowa. The Cedar Rapids Public Library has partnered with The Gazette to make millions of pages of the newspaper available online. The Gazette dates back to 1883, and the new database is keyword searchable. A recent article reports that 2 million pages are currently available online in this searchable archive, with plans to digitize another 1 million pages over the next 18 months.

Kansas. From a recent article: “Complete issues of Fort Hays State University’s Reveille yearbooks – from the first in 1914 to the last in 2003 – are now online, freely available to the public in clean, crisp, fast-loading and searchable digital versions in Forsyth Library’s FHSU Scholars Repository.” Click here to go directly to the yearbook archive and start exploring.

Maryland. New at Ancestry.com: Maryland, Catholic Families, 1753-1851 (a small collection of 13.5k records, but an important point of origin for many US families). “Judging from the 12,000-name index at the back of the volume, for sheer coverage this must be the starting point for Western Maryland Catholic genealogy,” states the description for this collection of birth, baptismal, marriage, and death records for the parishes of St. Ignatius in Mt. Savage, and St. Mary’s in Cumberland, Maryland. Find a brief history of Catholicism in western Maryland with lists of priests and a summary of congregational growth. Then find lists of marriages, baptisms, deaths, and burials, and even lists of  those “who appeared at Easter Confession, confirmation, communion, or who pledged financial support for the parish priest.”

New Jersey. Findmypast.com subscribers may now access small but historically and genealogically important collections of baptismal records (1746-1795) and additional church records (1747-1794) for Hannover, Morris County, New Jersey. States the first collection description, “Despite being small in population, the township is rich in history. It was the first settlement established in northwest New Jersey, dating back to 1685, and is situated by the Whippany River.” The second group of records “pertains to an active time in Hanover, with the resurgence of religious revivals kicking off around 1740. The most populous denominations in the latter half of the 1700s were Presbyterian, Society of Friends (Quaker), Dutch Reformed, Baptist, and Episcopal.”

Pennsylvania. The Carlisle Indian Industrial School, located in Carlisle, PA, was a federally-funded boarding school for Native American children from 1879 through 1918. The Carlisle Indian School Digital Resource Center is a project that is building an online searchable database of resources to preserve the history of the school and the students who attended there.

They recently announced a new resource titled Cemetery Information. According to the site, this collection provides “easy access to a wide range of primary source documents about the cemetery and the Carlisle Indian School students interred there.” Available materials include an individual page for every person interred there with their basic information, downloadable primary source materials about their death, an interactive aerial map of the cemetery, and more.

Texas. The Texas State Library and Archives Commission has digitized a series of collections featuring archival holdings from the First World War through the Texas Digital Archive. These collections are:

  • The Frank S. Tillman Collection: “The bulk of the collection focuses on the Thirty-Sixth Division and also features items from the Ninetieth Division, the Adjutant General of Texas, and other Texas soldiers.”
  • General John A. Hulen Papers:”Highlights include correspondence, photographs, and scrapbooks, dating 1887-1960.”
  • 36th Division Association Papers: “The papers include correspondence, reports, military records, and scrapbooks, dating 1857-1954. Records relate to Texans’ experience during World War I, railroads in Texas, and the San Jacinto Monument.”

genealogy giants quick reference guide cheat sheetWhat genealogy websites are you using? Which additional ones should you also be using?

Learn more about the giant genealogy websites mentioned in this post–and how they stack up to the other big sites–in our unique, must-have quick reference guide, Genealogy Giants, Comparing the 4 Major Websites, by Genealogy Gems editor Sunny Morton. You’ll learn how knowing the relative strengths and weaknesses of Ancestry.com, FamilySearch.org, Findmypast.com and MyHeritage.com can help your research. There’s more than one site out there–and you should be using as many of them as possible. The guide does share information about how to access library editions of these websites for free. This inexpensive guide is worth every penny–and may very well help you save money.

Disclosure: This post contains affiliate links and Genealogy Gems will be compensated if you make a purchase after clicking on these links (at no additional cost to you). Thank you for supporting Genealogy Gems!

Pin It on Pinterest

MENU