All About GEDCOM

The GEDCOM digital file format is essential to genealogy. My expert guest from FamilySearch explains what a GEDCOM is, how to use it, and the most recent changes. He’ll also answer some of the most common GEDCOM questions. 

Show Notes

If you’ve been watching my videos for a while, then you probably know that I really recommend that you have a complete copy of your family tree on your own computer. But what if you’ve been building your family tree totally online up to this point?

The good news is that you can export your family tree as a GEDCOM file. But what exactly is a GEDCOM file?

Gordon Clarke,  the GEDCOM Developer Relations Manager at the free genealogy website FamilySearch.org joins me to answer that question and provide the latest information about the GEDCOM.

What is a GEDCOM?

(00:54) Lisa: What is a GEDCOM?

(01:14) Gordon: GEDCOM is actually an acronym for:

GEnealogical
Data
COMmunication.

It’s a type of file with specific rules that allows digital family history products to exchange information. It’s been around so long that all the software companies can read and export it.

Say for example that you have a particular family tree program you’ve been working in but there are some features in another application that you like to try out. You want to try it out with a computer file that the program can read. All of the popular genealogy programs allows you to write a GEDCOM file and then you can read it in and review your information and add to it. That is what a GEDCOM is for.

It’s a specific file type that was works with most family history applications. It’s a text-based file, though it has special constraints to it. It was designed to be easily adaptable and compatible with importing and exporting. So, as long as the developers of both products adhere to GEDCOM specifications, you shouldn’t have a trouble downloading from one and uploading to the other.

You can learn more at GEDCOM.info.

What GEDCOM stands for

Lisa: It sounds like each genealogy software database and website probably have their own proprietary file type, right? So, this is one everybody sort of agrees on that can extract the genealogy data set right. Is that right?

Gordon: Right, and there are differences between the proprietary program and GEDCOM. There are some products out there that only support GEDCOM. So that’s their proprietary format.

Why Use a GEDCOM?

(03:45) Lisa: So why should we use one a GEDCOM. When would we find ourselves wishing we had this universal file?

Gordon: Family history is more of a record keeping whether it’s photos and stories and genealogical data. People like to keep it and have control over it. So, GEDCOM is I like the word “personal”. You can personally control it. It’s just a .GED file, so any operating system can handle copying and emailing it. So, for personal control, preservation and sharing of genealogical data. It’s the most universally accepted format.

I would think for your backup purposes because it’s so universal, make sure that the program that you’re using has the ability to save your data in GEDCOM. Then you can decide whether you put it in your thumb drive or removable drive or you put it up in the cloud, you can decide how to preserve it. Think of it more as your personal file over this important information.

(05:31) Lisa: I like that idea. I’m probably not alone in that I once had somebody give me a little floppy disk and it had the whole family tree that this person had been working on. Unfortunately, it was a proprietary file, and it was a program that no longer exists. I’m helpless to be able to use it. So, a GEDCOME can really solve that issue.

Do All Family Tree Programs Support GEDCOM?

(06:00) You kind of touched on this, but I just want to just double check. Can all family tree programs and websites export the GEDCOM? Are you familiar with anything that don’t?

Gordon: I would say all of the popular programs and websites make it possible to import GEDCOM, and most of them allow for exports. There are some exceptions to that rule. If you’re going to spend your time using a program, look to see if it’s GEDCOM compatible.

To help even more so standardize the industry, the software providers commit to implementing the newest version of GEDCOM. Much of that is backward compatible. We presented those that have or will be planning to implement the newest version of GEDCOM at Rootstech. You can search at Rootstech for “GEDCOM” and see the videos of what’s been rolled out and what’s coming.

Who Owns and Controls GEDCOM?

(07:41) Lisa: Is there one particular group or authority or somebody who’s in charge of deciding what the GEDCOM is and how it works? Or is that a role that FamilySearch is playing?

Gordon: It is a role that Family Search has been playing. FamilySearch is the software development, education marketing, support arm of the department called The Family History Department of The Church of Jesus Christ of Latter Day Saints. So sometimes because of marketing reasons, people think that we’re different. Family Search is totally run by the Church of Jesus Christ of Latter Day Saints.

From a historical standpoint, the original specification was created and released in 1984. All subsequent versions have been copyrighted by the Church of Jesus Christ of Latter Day Saints.

Now, in the last three years, as a like a product manager, I took on the responsibility for working on the new version, version 7 of GEDCOM. But it’s always been an effort of FamilySearch as the outreach arm for the Family History Department.

What we did differently in this last version is we solicited all the key players and software companies. It was much more of a collaborative effort to go through the changes, things to keep, things to just get rid of. It took about two years working with many people. Now the version is what is called a public GitHub repository. As we worked toward version 7, it was to prepare it for a starting point. The decision process is still a steering committee sponsored by FamilySearch, but the input and the communication on changes is open to all software developers. You can learn about all that because it’s hosted at GEDCOM.io. So GEDCOME.info is kind of like the general public, and GEDCOM.io is more for technical software developers.

GEDCOM Features

(10:38) Lisa: What are some of the features of GEDCOM 7? What are some of the things that you consider when you’re continuing to develop the GEDCOM?

Gordon: The process that we worked on was, I think to eliminate ambiguity, there could be different software providers that would interpret the file specifications a little bit differently. We wanted to clean up the specifications so that there would be much more, not 100%, but a much better compatibility between the people that were reading it and writing it. So, we worked very tediously on eliminating the ambiguity.

I would think that the biggest thing is, it’s become more of a storage format of photos, and records and data. Let me read something, “FamilySearch GEDCOM version 7 incorporates the added ability to include photos, and other files when users download a FamilySearch GEDCOM 7 file from a supportive family tree product.”

Your local photos can be bundled in a special file that we called GEDZIP. It’s a GEDCOM file that is a zip package. That means that anybody that unzips that package will get the GEDCOM file and all the external files associated with it and have everything be readable. It’s a packaging technique to put everything together, which really adds to this idea of a personal preservation and sharing. Now you can package everything together and preserve it and share everything that’s important to you with others.

In addition to this zipped packaging capability, notes have been expanded for more versatile use and styling of text. When you add notes, whether it’s a relationship or a location, you can actually stylize those notes now and use bold and italic.

Many tools and sample files were created to help with self-testing. It’s based upon the Apache license, which is more of a technical slant on things, but to software developers, that means it’s an open software license. There’s a public GitHub repository that you go to github.com/familysearch so that you can request and watch ongoing changes in a more of a public environment, though Family Search is still the stewards and has the final say on decisions.

So that’s what’s new. It’s more open to the public. It’s been cleaned up with some important new features.

But backward compatibility for 90% of the GEDCOMs that are out there (and the last one was 5.5.1) is still possible. But it won’t go back to 3.0, 2.0. That’s where that’s where some of the incompatibilities are, is because people are using versions that are 20 years old. And things have changed a lot in the last 25 years. We have a clean, fresh start, and a new community working on continuous improvements. But there won’t be changes because the standards shouldn’t change much. This new version 7 is going to be pretty much the same for a while everybody gets on board.

Do GEDCOMs Include Image Files of Attached Records?

(15:21) Lisa: You mentioned photographs. Would that include image files? Would that include if we downloaded an image of a genealogical record which might be a .JPEG file? Would those come along with the GEDCOM?

Gordon: Yes, absolutely. All the elements of GEDCOM have definition of how to use them. And what’s called the multimedia link, the multimedia link means you can link to local files, JPEGs, PDFs, you know, whatever they are. And if you don’t want to put it all together, you can link to files that are in the cloud, and it will remember where they are. If you package them together in a GEDZIP file, and then you unpackage it, you’ll be able to access the local image files and the local records there.

So, this idea of putting it all together, I mean, bandwidth is much better than it used to be. But still, for people that have hundreds of thousands of images. This is not the best format for that. So they can work out a strategy taking into account the cloud service they use, and which photos they will keep locally on their computer. So, they can keep track of everything, both in the cloud and on their local drive. And that can all be referenced in this new version of GEDCOM.

Is There Data Loss When Exporting a GEDCOM?

(16:59) Lisa: Excellent.

So one of the questions I’ve heard from people is that they are concerned about loss data loss. If they’re importing or exporting, maybe going back and forth, is there a chance that you’re going to lose things or even introduce an error of some type?

Gordon: This is kind of the issue of the work on version 7. One of the biggest issues is not only new features, but to get a new standard to kind of clean the slate. If you get stuff into the new GEDCOM version 7 the likelihood of data losses is greatly reduced. So, we’re encouraging the adoption and use of GEDCOM 7 because it’s less likely to cause any data loss or errors.

Family Search and industry experts have worked for two years to remove ambiguities, simplify the definitions and samples in order to eliminate the possibility of data loss and errors when transferring between programs. In the long run, not only does it include more media, but the whole goal is to improve the consistency, the compatibility and minimize or even eliminate data loss. So, what you will start being seeing is the question “is GEDCOM 7 compatible?” Because GEDCOM 7, when we were working on something that was 20 years old, is going to be more compatible in the future. We have a body to watch out for it. Your data will migrate to the new version without data loss. But looking at down the road, staying with the version 7 or higher will assure a sure better preservation of what you have.

Learn More About GEDCOM at Rootstech

(19:17) Lisa: I think you mentioned or alluded to that there were some announcements at Rootstech 2022.

Gordon: Yes, go to the sessions and type in “GEDCOM” and you will get three opportunities. One is a session called GEDCOM 7 Launched and Rolling Strong. Another session will be FamilySearch GEDCOM 7 What’s Next? And the answer is teamwork.

There’s two pre-recorded videos about the What’s New in GEDCOM 7 and then how the industry’s going to join together in working on it in the future. In in one of the sessions, the first one, there actually is a slide that shows all the companies that have committed to it. But all the majority of the companies have said, both in the cloud and desktop and laptop, and some have said when they’re going to release it. And one company I think, is announcing their release at Rootstech of the new GEDCOM version 7.

Future Updates and Changes to GEDCOM

(20:44) Lisa: That’s great to see. Anything I didn’t ask you or that you think people should really be aware of as they move forwarding and keeping up to date with GEDCOM 7?

Gordon: Again, with a standard, we don’t want to change too much too fast, because they wanted to get solid as a new transfer format.

I think the big areas that we’re working on for future versions is related quite a bit to internationalization. There are probably 20 different calendaring systems that are different than what we do in the U.S. To be able to respect those different calendars and to understand the translation between calendars is a big part of internationalizing GEDCOM.

The other part related to that is that there are some places in the world where how they define relationships between people is not typical to either the US or Western Europe. And so we are working on major upgrades and encourage people to come join with us. With naming conventions we may think given name, surname, but in reality, there’s other relationships that get into the name. If we even go to Africa their name is the first name may go back 10 generations, so their name is a memorization of all those names. So, improving on names is an important effort, the structure and relationships.

Another improvement is places. We think hierarchal and certain jurisdictions, but over time, and in different areas of the world, how you organize places is different. We need to address that in the GEDCOM specification.

Sources and Citations need to be upgraded for the genealogical community. And so, we certainly invite not only software developers, but genealogists to join our effort to improve sources and citations.

GEDCOM Hypothesis

One thing I’m really excited about is that we have a team that’s been working a year, and they’re probably working on it another year or two, on what we call hypothesis. This is so that you can share information without claiming it as a conclusion, and keep it separate from a conclusion. This encourages collaboration. So instead of arguing about I’m right, you’re wrong, we call it a hypothesis. Then we can have a discussion until there’s enough sources to prove it. This Hypothesis module I think is going to be really exciting. But that won’t be for a couple years or so until we actually release it.

Lisa: I think that’s a terrific idea because so often we are just battling with ourselves over what we think the answer is, and we want to track it while we’re doing it.

I’m curious: sometimes we go to a website, and you have to pick what language you speak. Perhaps if you’re searching for videos on YouTube you might say English. Is this something being considered? Is the goal no matter what that it’s only one type of file that serves every country or was there a consideration that you could select your country and then the GEDCOM would support your calendar and your geographic areas. I’m sure that was a discussion.

Gordon: Oh, absolutely. And, but what you’re talking about, just to be clear, is the specification to give all of the options and more to the software developer. The software developer can decide the language of the interface, and many of them are already doing this. So the actual presentation, if it’s Norwegian, or Danish, or whatever, it’s different according to the language that you place. What we’re looking according to your language of choice is that the orientations are names, relationships, and places jurisdictions, will be easy for the software developer to switch to by just changing that.

When we look at an international – how people look at information – it may be a different lens that they look through. So having the ability to give the software developers out of our future specs, to switch their interface, and switch around because they might be working in one part of the country because of their heritage, and then they might work in another and to be switched between it and to still have the data be the same, regardless of what national lens they’re looking through.

Lisa: It’s amazing that one little package contains so much and so much flexibility. That’s really terrific.

The Team Working on GEDCOM 7

(26:52) Gordon: I won’t drop names but in my immediate steering committee, that we meet with weekly, not only do I have three representations from within FamilySearch, but from the community, I like to call them doctors, they are doctors, they have their PhDs in computer science. Some are genealogists, they have their peers, one is even a linguistic professor. Another is an actual legal professional. It’s been wonderful to work with such experts, really, that are reasonable, and want to make things easy for the software developer. So, it’s quite a dilemma, instead of just making it right in the specification, but we’ve got to make it right and make it easier for the software developers to implement it. So that’s my thanks to all the people I’ve been able to work with.

GEDCOM Resources

(28:19) Lisa: Visit GEDCOM.io and GEDCOM.info.

Are they able to offer any volunteering opportunities? Do you need the help of people who are doing genealogy?

Gordon: Oh yes, you can volunteer in lots of different ways at GEDCOM.io.

Lisa: Thank you so much for taking time to explain GEDCOM.

Resources

Downloadable ad-free Show Notes handout for Premium Members. (Learn more and join Premium Membership here.)

 

A Tip for Harnessing New Technologies for Genealogy

Lisa BYU Keynote

Photo courtesy of The Ancestry Insider

New technologies don’t stay new. They keep evolving. Here’s a tip for harnessing new and emerging technologies to advance family history research and stay connected with living relatives. 

Last week, I was at the BYU Conference on Family History & Genealogy in Provo, Utah. What a friendly, welcoming group! (Be sure to check out the BYU Family History Library here.) All week, I taught sessions and gave a keynote address on various technologies that help our research. The week’s discussions reminded me how quickly technology moves–and how enthusiastically genealogists continue to embrace new opportunities given them by technology.

It’s part of my job to learn about these new technologies and pass the best ones–the “gems” along to you. But here’s a tip I shared during my keynote address that will help you focus on the technologies you care most about: Think about which tasks you want to accomplish with technology, rather than just learning genealogy-specific technology. Then keep up with developments in the technologies that accomplish those tasks.

For example, by now, many of us have used (or at least heard of) Google Translate. We can use it with foreign-language documents and to correspond with overseas relatives and archives. But Google Translate’s functionality keeps improving. “By the audible gasps of the audience” (during my keynote address) reported the FamilySearch blog, “most were not aware that the Google Translate app enables you to literally hold up your phone to the computer screen or typeset document, and it will translate foreign text on the fly for you—a must have free tool when dabbling in nonnative language content.”

Genealogists are really thinking about these issues. The Ancestry Insider blogged about my keynote talk, too, and my observation that genealogists haven’t been embracing digital video at the same speed at which they embrace other forms of digital media. In the comments section of that post Cathy added, “Now what we need to do is get FamilySearch to figure out a way to let us upload our URL YOUTube videos, not only for our deceased, but for our living….Our children and grandchildren don’t write letters, they email, text, instagram. They don’t write journals, they blog. They make videos of current history….We all need to look to the future and [learn] how to save the new technologies.” Cathy gets it!

A special thanks to conference organizers Stephen Young and John Best, who welcomed me and Genealogy Gems Contributing Editor Sunny Morton all week long. They did a fantastic job of organizing a large event while retaining a warm, personal environment.

Continue reading about applying technology to your family history here.

Prison Inmate Photos: “The Eyes Are Everything”

Matt from Omaha, Nebraska (U.S.) recently told me about a project his cousin is working on that is so cool the story was picked up by U.S.A. Today.

Prison Memory

While poking around at an 1800s-era Iowa prison about to be torn down, Mark Fullenkamp came across boxes of old glass negatives. Upon closer inspection, he found they were intake photos of the inmates. Some were 150 years old!

Mark first set out to digitize and reverse the negative images of over 11,000 prison inmate photos. Others gradually became involved, like scholars at University of Iowa where he works and even inmates at the Iowa Correctional Institution for Women. A doctoral candidate who was interviewed by U.S.A. Today says she’s struck by the moment these photos were taken: when their lives were about to change forever. Though many look tough for the camera (and presumably the other inmates), she sees a lot of emotion in their expressions: “The eyes are everything.”

Now Fullekamp’s team is trying to connect names and stories with the photos. It’s not easy, but many of the pictures have inmate numbers on them. Some files have surfaced with inmate numbers and names in them. Others are stepping forward with memories.

Read more about the project on Matt’s blog.

Got a digital photo archiving project of your own? Click here to learn about a free ebook published by the Library of Congress on digital archiving.

We Dig These Gems! New Genealogy Records Online

We dig these gems new genealogy records online

Every Friday, we blog about new genealogy records online. Do any of the collections below relate to your family history? Look below for early Australian settlers, Canadian military and vital records, the 1925 Iowa State Census and a fascinating collection of old New York City photographs.

AUSTRALIAN CONVICT RECORDS. Now Findmypast subscribers can access several collections on early settlers. Among them over 188,000 Australia Convict ships 1786-1849 records, which date to “the ships of First Fleet and include the details of some of the earliest convict settlers in New South Wales.” You’ll also find “nearly 27,000 records, the Australia Convict Conditional and Absolute Pardons 1791-1867 list the details of convicts pardoned by the governor of New South Wales and date back to the earliest days of the colony” and New South Wales Registers of Convicts’ Applications to Marry 1825-1851, with over 26,000 records.

CANADIAN WWI MILITARY RECORDS. As of June 15,  162,570 of 640,000 files are available online via the Soldiers of the First World War: 1914–1918 database on the Library and Archives Canada website. This is the first installment of an ongoing effort to digitize and place online records of the Canadian Expeditionary Force service files.

IOWA STATE CENSUS. About 5.5 million newly-added records from the 1925 state census of Iowa are now free to search at FamilySearch,org. Name, residence, gender, age and marital status are indexed. The linked images may also reveal parents’ birthplaces, owners of a home or farm and name of head of household.

NEW YORK CITY PHOTOGRAPHS. About 16,000 photos of old New York City from the New York Historical Society are free to view on Digital Culture of Metropolitan New York. According to the site, “The extensive photograph collections at the New-York Historical Society are particularly strong in portraits and documentary images of New York-area buildings and street scenes from 1839 to 1945, although contemporary photography continues to be collected.”

ONTARIO, CANADA VITAL RECORDS. Nearly a half million birth record images (1869-1912), nearly a million death record images (1939-1947) and over a million marriage record images (1869-1927) have been added to online, indexed collections at FamilySearch.

check_mark_circle_400_wht_14064Today’s list of new records has a LOT of Canadian material! If you’re researching Canadian roots, here’s a FREE video for you to watch on our YouTube channel: Lisa Louise Cooke’s interview with Canadian research expert Dave Obee, who shares 10 tips in his effort to help one RootsTech attendee break through her brick wall. This post and tip and brought to you by The Genealogist’s Google Toolbox by Lisa Louise Cooke, newly-revised and completely updated for 2015 with everything you need to find your ancestors with Google’s powerful, free online tools.

Pin It on Pinterest

MENU