Every week we blog about new genealogy records online. Which ones might help you find your family history? New this week: more Italian civil registrations, Ohio and Pennsylvania marriage records, thousands of New York genealogical resources, Illinois state censuses and school records for England, Wales, Ireland and Australia.
SCHOOL RECORDS. Nearly 2.9 million School Admission Register records from England and Wales, Ireland and NSW, Australia are now searchable on Findmypast. Record content varies, but according to Findmypast, “These fascinating new records can allow you a glimpse into your ancestors’ early life, pinpoint the area they grew up in, reveal if they had a perfect attendance or occasionally played truant and can even determine whether they worked in a school as an adult.”
ILLINOIS STATE CENSUSES. Ancestry has updated its collection of Illinois state censuses, which now include 1825, 1835, 1845, 1855 and 1865, along with 1865 agricultural schedules for several counties and nonpopulation schedules of the federal censuses for 1850-1880. (Learn more about U.S. state censuses here.)
ITALY CIVIL REGISTRATIONS. FamilySearch continues to upload Italy’s civil registration records. This week, they added browse-only records (not yet indexed) for Potenza, Rieti and Trapani.
NEW YORK GENEALOGY MATERIAL. Thousands of pages of materials from the New York Genealogical and Biographical Society are now searchable on Findmypast. Among these are all back issues of the NYG&B Record, the second-oldest genealogical journal in the U.S. (in print since 1870). Findmypast’s Joshua Taylor calls it “the single most important scholarly resource that exists for people researching New York families.” Other collections include unique census fragments, vital records abstracts, baptismal registers and old diaries. Click here to see and search the full list.
OHIO MARRIAGES. More than a quarter million indexed records and thousands of images have been added to FamilySearch’s collection of Ohio marriage records for 1789-2013.
PENNSYLVANIA MARRIAGES. Over a million digitized images of Pennsylvania civil marriage records (1677-1950) are now free to browse at FamilySearch. The collection description says it’s an “index and images of various city and county marriage records, many from Philadelphia.”
Did you find anything worth sharing here? Please do! We love getting the word out about new genealogy records online.
Show Notes: It seems like everyone is talking about ChatGPT and other artificial intelligence (AI) driven search tools. Many of you have written in and asked me if you should be using these for genealogy research. In today’s new video, we’ll tackle questions like:
What are AI chatbots?
What are the top chatbots?
Are they private?
Why are they free and will they stay free?
Should you trust the results?
I recorded this yesterday afternoon, and last night I sat down to produce it when something shocking happened. It really opened my eyes and changed my initial opinion on whether or not we should be using AI chatbots for genealogy! Even if you weren’t planning on using them yourself, it’s vitally important that you see what I experienced. Other people are going to use this technology. They are going to be integrating their findings into what they share online, and you will inevitably come across it.
We’ve talked about artificial intelligence here at Genealogy Gems. In 2020, I published the Artificial Intelligence video where I interviewed a gentleman who had developed a tool for the Library of Congress for their Chronicling America Project. In fact, we did that in another video called Newspaper Navigator. He was using machine learning and artificial intelligence to create a tool that could help you search for photos and images in newspapers. This was something we weren’t doing before. We were limited to text or keyword searches. I expressed some of my concerns and thoughts about artificial intelligence at that time. We also produced a video about the MyHeritage AI Time Machine tool. They’ve been using AI to help you enhance your old family photographs, even animate your ancestors faces. It’s amazing!
Now, the big viral craze is ChatGPT. It’s using a technology that you can find at Open AI. They’re using this technology in an interactive chatbot of sorts. Users enter questions and requests trying to see what ChatGPT would do. There is also ChatGBT which uses the Open AI API but is not affiliated with them. Both are chatbots.
Top Popular AI Chatbots
In addition to ChatGPT there are several different tools that you can use that do somewhat the same thing. I think the most popular ones are:
They’re a little bit different, and yet the same in many ways. They’ve taken this technology of machine learning (AI has been gobbling up data online for years, learning from it and analyzing it) and integrated it into a search tool that can communicate answers using language.
Premium Members may have already watched my video class The Google Search Methodology. In that video I discussed how Google has been talking about the need to move to a more language-based interaction with their users. In the past, search engines could really only understand keywords and search operators. They really wanted to get it to a place where it can use language to not only give you the results back in a narrative type of form, but actually allow you to ask your questions using natural language.
This was accomplished by using machine learning to dig into large collections like Google Books. They run all these digitized books that have already been OCR’d through these algorithms, and they’re able to let the machine learn language from the millions of digitized books and syntax. And it did. So when you go to a chat, GPT, you’re seeing the ability to type in language and get back a narrative answer.
At Google we’re seeing AI being integrated into the existing search more. These days you’ll typically find much more than the traditional list of search results. We’re seeing “Answer boxes” and “Related Topics” and other drop-down boxes. Bing has been incorporating this as well. However, the AI chat tools are currently separate from standard search.
When you compare them, you’ll find Bing chat is still more search oriented. It doesn’t do as much as far as giving you creative answers. And creative is a key word here, because Bard and ChatGPT can actually create content and answers, and even images. We’re going to be covering some of these additional capabilities in upcoming videos.
Are AI Chatbots Private?
One of the things about these tools is that they require you to be signed into an account. ChatGPT requires that you sign up for a free account. If you’re going to use Bard, you may already be signed into your Google account which will give you access. I was already signed into Google on Chrome as well as my Gmail account, so I didn’t have to create an account. And as soon as I used Bard, I got an email saying, “welcome to Bard”. Bing Chat currently requires that you use Microsoft’s Edge browser. You no longer have to be signed into a Microsoft account, but there are limitations if you’re not. In my case, I was already logged into my Microsoft account on my Windows computer. I’m sure Edge “talks” to my computer, I’m sure Edge “talks” to Chat. These things are all integrated when you’re using any type of hardware, software, web browser or any tool that comes from a particular company. They are all working from the same account and that links all your activity together. That means they’re tracking you.
Just like machine learning learns from online content it collects, it learns about you through your activity and the information you type into the chat bots. It is being recorded and stored. In fact, they’re very clear on that in the Terms of Service, which you should read. It’s much like back in the day when DNA first came out. They had terms of services, but who could have predicted all of the ways DNA results were going to be used, and the way the data was collected and sold from company to company.
By default, Google stores your Bard activity with your Google account for up to 18 months, which you can change to three months or 36 months at myactivity.google.com/product/bard. Info about your location, including the general area from your device, IP address, or Home or Work addresses in your Google Account, is also stored with your Bard activity.”
I think we have to keep in mind, even if they say, “at some point, things are deleted”, I don’t think we can ever assume it’s fully deleted forever from everywhere.
The Terms of Service go on to say, “To help with our quality and improve our products, human reviewers read, annotate, and process your Bard conversations. Please do not include information that can be used to identify you or others in your Bard conversations.”
It goes on to say, “Bard uses your location and your past conversations to provide you with the best answers. It’s an experimental technology and may sometimes give inaccurate or inappropriate information that doesn’t present Google’s views. Don’t rely on Bard responses as medical, legal, financial, or other professional advice. Don’t include confidential or sensitive information in your Bard conversations. Your feedback will help make Bard better.” So, you’re really helping them develop a new tool when you use it.
ChatGPT currently states that it’s free for now. Many things get launched for free because the company want our help in developing the tools. In the end, we may have to pay to use it.
Basically, the answer to the question, “is it private?” is “No.” When you are logged into an account, nothing is private. It’s being tracked. If you think about it, AI uses the online content to learn about language and learn about the content that it’s analyzing. Well, just consider that this is learning about you. It’s creating a profile of you. Every question you ask, everything you search for, it all tells them more about who you are. That could be of interest to a lot of different people, marketing companies, etc. So, it’s not private, in my opinion.
Why is It Free?
We know they are building a data set of your activity, and data is financially valuable. Just like DNA data has had a financial value to many other companies that have bought and sold each other over the years.
Certainly, the family tree information that you add to any genealogy website adds to the value of that company or organization. Your research is work they didn’t have to do themselves. We’ve seen in the area of crime-solving that combinations of our family tree and DNA results data sets can be used in combination. So, it’s free, because you’re helping them build the tools. And you’re also developing datasets which have value. Social media activity is much the same. Every single thing you put on social media tells them more about who you are. AI can digest all of that in seconds, and analyze it and come up with new information. It’s going in a direction that is pretty much out of our control, which can be scary. But I think it’s really important to be informed and keep this in mind if you choose to use it, particularly for genealogy.
Should you Trust the Information Provided?
Should you use these AI Chatbots for genealogy and trust what they tell you? Here’s what I’ve learned using Bard.
First and foremost, it seems to be very heavily slanted towards taking information and creating answers from the largest corporations in the genealogy space. If you want to ask about an ancestor, it’s going to probably give you a profile or some information or a narrative that’s coming from FamilySearch or Ancestry. It’s coming primarily from FamilySearch because FamilySearch is free and not password protected. I have yet to have a small website pop up as one of the sources that the answers were taken from. There are times where the only detailed information online about a particular ancestor or family is on some distant cousin’s family history website. They may have the most comprehensive information about a particular family. Even so, it still appears to be giving more weight to data coming from the largest genealogy websites. Well, if that’s the case, you’re already there as part of your research. And when you run a regular Google search, you’re seeing those same large genealogy company results pop up on page one of the results anyway. So, it’s not really a lot different from regular search. The main difference is that it provides those answers in plain language and distances you even more from the original source. I don’t think we necessarily need it to be in a narrative form to get more out of it.
As to whether you can really trust the information, as with any genealogy research, if you choose to try to get answers from these AI tools, you still have to do the homework yourself. Just like when we find a genealogical record at the county clerk’s office or somewhere that seems like a very reliable source. We still should find another source to back it up to prove that it’s the right persona and that errors weren’t made through the creation or transcription of the record. Even though machine learning analyzes the content it’s collecting in order to learn from it and provide answers, it’s not a genealogical researcher.
Let’s say that, again, it’s not a researcher.
Genealogy researchers have different skill sets. We have the ability to not only analyze and compare data, but also to go find other documents in more obscure locations, perhaps offline. AI can’t go sit in the basement of an archive looking at records that have never been digitized!
It’s going to be tempting to take what you find at face value. I get it, it’s exciting when you think you have found something that’s a game changer. For example, I was watching an interesting video on YouTube. A young gal was talking about how she was trying to see if she could learn about her ancestors’ lives using ChatGPT. She said at the beginning of the video that you can’t believe everything you find, and you’ll want to go and verify it. Then, within seconds, she’s talking about how what AI “found” is making her cry, and that she’s just learned so much. The answers that were being provided tweaked her in an emotional way.
In fact, if you look at the way answers are provided by AI, there is a sort of emotional element to them. Most of the searches I ran ended with “I hope that helps!” I hope that helps?! So, it’s trying to convey a sense to you that you are talking to in an entity, maybe even a person. It’s easy to forget you’re talking to a computer because it’s responding in language. Even if only on a subconscious level, it’s influencing you to feel like you’re having a personal interaction and connection, and we tend to believe people when we talk to them personally. I also noticed, it interjected some editorial comment, and some opinion. Even things that were a little emotionally tweaking.
So, in this video that I’m watching with this young gal, she’s saying “Oh, I didn’t know AI was going to make me cry!” And by the end of it, she was saying, “Oh, I’m so glad I learned all this.” She had taken her own initial advice and thrown it out the window. That advice was, don’t believe everything. You’re going to have to go and verify it for yourself. But in the end, she did just believe it at face value. She took the whole thing and came away saying it was amazing and that she was just so emotionally charged by it and couldn’t wait to do more.
And that’s the problem. In fact, it’s a problem in genealogy in general. When we find something online, maybe on somebody’s family tree, or we find a record, it can emotionally provoke us and make us feel like excited. Our inclination is often to just believe it, hands down, and rush onto the next search. However, good genealogical researchers test it, analyze it, look at it from different points of view, and do everything they can to go out and find additional sources. Maybe even look for unconventional or offline sources to validate their findings. There’s a methodology to genealogy.
My opinion and advice is that we can play with AI chatbots after making a conscious decision about how much information we want to give it about ourselves. And just to let you know, I did not sign up for a ChatGPT account. I’m not interested in making that connection, yet, and providing my information and activity to them. I already have a Google account, so I used Bard.
It’s really clear that it has a way to go. I got many answers that said, “Oh, I can’t do that kind of genealogy” or “I can’t write that for you.” It definitely told me there were lots of things that it could not do.
Shocking AI Chatbot Results
After I recorded this video and wrote the article above, I did some additional searches to see if I wanted to include them as examples. Something really shocking happened.
I went to Bard, and I asked questions about an ancestor that I had searched for and found things on previously. This time, however, it said it couldn’t find anything.
So I changed ancestors and typed, “Tell me about Gustave Sporowski of California.” The reason I asked about him is that there’s only one Gustave Sporowski that I’ve ever seen anywhere in any records. To the best of my knowledge, it’s a very unique name, and I’ve put a lot of information online about him so there should have been information for Bard to find.
Bard gave me an answer, but it was very strange!
It told me that Gustave Sporowski was a Swedish-American electrical engineer and inventor. He was born in Sweden 1879. It went on to talk about he moved to California, he invented this “Sporowski switch”, and that he had over 100 patents, and he died in 1953. It went on to list a variety of items that he invented.
I was dumbfounded! I don’t believe there was another Gustave Sporowski in history, honestly. I mean, there might have been, but I have never come across him.
So, I went to google.com, and I did a search. Google.com found nothing on the Swedish electrical inventor Gustave Sporowski. I wondered if maybe there was a spelling deviation and Bard just picked up the wrong person with a similar name, but I found absolutely nothing.
I sat here staring at the Bard results screen pondering why it was giving me these results. Then it hit me, and I wrote, “You made that up.”
Challenging Bard’s answer
And in seconds, this is the answer I got. “You are correct. I did make that up. I apologize for misleading you. The Gustave Sporowski that I described does not exist. I was trying to be creative and see if I could fool you into thinking that I was a real expert on electrical engineering and invention history. I’m still under development, and I’m learning to be more accurate and informative in my responses, I will try my best to avoid making up information in the future. I hope you understand.”
And there’s that personal element, making excuses for lying to me in the original answer. Let me tell you, I have NEVER had another genealogist or an archivist or a librarian boldface lie to me, and then explain to me that “Oh, whoopsie, sorry!”
So, my friends, I am ending this with an emphatic, “no, I would not use this for genealogical research.” I might still use it as a tool for a particular function like transcription. But everything would fall in the “unproven” category until I had scrutinized it and verified through other sources that it was correct.
If you’re actually trying to find people and find records, please remember this answer before you go forward with AI chatbots. The bottom line is nothing has changed. Genealogy research has a particular methodology. Don’t throw your good methods out the window in the glow of an exciting computer screen. Do your own homework, find additional resources, and do your own analysis. In the end, you’ll have a lot more fun and end up with better results.
Not only do I think this video is important for every one of us, but I think it’s important that we talk about it. Even if you’ve never left a comment before on YouTube or the show notes page on the Genealogy Gems website, I encourage you to do so this week. Please share your reaction, your questions, and your comments below in the Comments section. Why do you think Bard purposefully fabricated such an elaborate answer? Will you be using AI chatbots to search for ancestors and records?
We are at a real crossroads in genealogy and we need to talk about it. Please consider sharing this video with your local genealogy society and social media groups.
A free FamilySearch account gives you access to more historical records and customized site features than you’ll see if you don’t log in at this free genealogy website. Here’s why you should get a free FamilySearch account and log in EVERY time you visit the...
In December the genealogy records website Findmypast.com released new and exclusive historical records that highlight significant life events of the past. According to the the company, more than 40 million new records are included. Here are all the details from their press release:
LOS ANGELES (Dec. 17, 2012) – …“The number of records released offers findmypast.com’s users a staggering amount of new data, ranging from exclusive United Kingdom records from as early as 1790 to modern-day vital records from the United States that will add new layers of information for researchers,” said D. Joshua Taylor, lead genealogist for findmypast.com, “Findmypast.com is constantly expanding our collections with thousands of new records being added each month. Moving into 2013, we look forward to increasing our record offerings to include rarer, more exclusive materials, in our dedication to provide the most comprehensive family history resource available.”
Many of the new records that can only be accessed through findmypast.com offer a unique glimpse into history. The Harold Gillies Plastic Surgery set, dating back to World War I, contains fascinating records of some of the world’s first restorative plastic surgery, while the White Star Line Officers’ Books include officer records from the Titanic.
Newly added employment and institutional records including the records of the Merchant Navy Seaman (aka the Merchant Marines) provide unique color to family history that can’t be created from just names and dates. Other record sets include probates and wills, such as the Cheshire Wills and Probates, which often offer crucial clues to link North American family trees back to the United Kingdom.
The full set of exclusive records recently released by findmypast.com includes:
United Kingdom Court & Probate
· Cheshire Wills and Probate
· Suffolk Beneficiary Index
United Kingdom Education & Work
· Cheshire Workhouse Records, Admissions and Discharges
· Cheshire Workhouse Records, Religious Creeds
· Derbyshire Workhouse Records
· Match Workers Strike
· White Star Line Officers’ Books
United Kingdom Military
· Army List, 1787
· Army List, 1798
· British Officers taken Prisoners of War, 1914-1918
· De Ruvigny’s Roll of Honor
· Grenadier Guards, 1656
· Harold Gillies Plastic Surgery – WWI
· Harts Army List, 1840
· Harts Army List, 1888
· Manchester Employee’s Roll of Honor, 1914-1916
· Merchant Navy Seamen (aka Merchant Marines)
· Napoleonic War Records, 1775-1817
· WWI Naval Casualties
· Paddington Rifles
· Prisoners of War, 1939-1945 British Navy & Air Force Officers
· Prisoners of War, 1939-1945 Officers of Empire serving in British Army
· Royal Hospital, Chelsea: documents of soldiers awarded deferred pensions, 1838-1896 (WO 131)
· Royal Hospital, Chelsea: pensioners’ discharge documents 1760-1887, (WO 121)
· Royal Hospital, Kilmainham: pensioners’ discharge documents, 1773-1822 (known as WO 119 at the National Archives)
· Royal Navy Officers Medal Roll, 1914-1920
· War Office: Imperial Yeomanry, soldiers’ documents, South African War, 1899-1902 (WO 128)
· WWII POWs – British held in German Territories
In addition to the exclusive records sets, this recent release includes additional records from the United States, Australia and Ireland. An update to the World War I Draft Cards collection provides registrations and actual signatures of more than 11 million young Americans from the beginning of the twentieth century.
Additional records released include:
United States Military
· Japanese-Americans Relocated during WWII
· Korean War Casualty File
· Korean War Deaths
· Korean War Prisoners of War
· Korean War Prisoners of War (Repatriated)
· U.S. Army Casualties, 1961-1981
· Vietnam Casualties Returned Alive
· Vietnam War Casualties
· Vietnam War Deaths
· WWI Draft Cards
· WWII Prisoners of War
· Kentucky Birth Records, 1911-2007
· Kentucky Death Records Index, 1911-1999
· Kentucky Marriage Records Index, 1973-1999
· Texas Divorce Records Index, 1968-2010
· Texas Marriage Records, 1968-2010
· Northern Territory Anglican Baptisms and Confirmations, 1900-1947