All About GEDCOM

The GEDCOM digital file format is essential to genealogy. My expert guest from FamilySearch explains what a GEDCOM is, how to use it, and the most recent changes. He’ll also answer some of the most common GEDCOM questions. 

Show Notes

If you’ve been watching my videos for a while, then you probably know that I really recommend that you have a complete copy of your family tree on your own computer. But what if you’ve been building your family tree totally online up to this point?

The good news is that you can export your family tree as a GEDCOM file. But what exactly is a GEDCOM file?

Gordon Clarke,  the GEDCOM Developer Relations Manager at the free genealogy website FamilySearch.org joins me to answer that question and provide the latest information about the GEDCOM.

What is a GEDCOM?

(00:54) Lisa: What is a GEDCOM?

(01:14) Gordon: GEDCOM is actually an acronym for:

GEnealogical
Data
COMmunication.

It’s a type of file with specific rules that allows digital family history products to exchange information. It’s been around so long that all the software companies can read and export it.

Say for example that you have a particular family tree program you’ve been working in but there are some features in another application that you like to try out. You want to try it out with a computer file that the program can read. All of the popular genealogy programs allows you to write a GEDCOM file and then you can read it in and review your information and add to it. That is what a GEDCOM is for.

It’s a specific file type that was works with most family history applications. It’s a text-based file, though it has special constraints to it. It was designed to be easily adaptable and compatible with importing and exporting. So, as long as the developers of both products adhere to GEDCOM specifications, you shouldn’t have a trouble downloading from one and uploading to the other.

You can learn more at GEDCOM.info.

What GEDCOM stands for

Lisa: It sounds like each genealogy software database and website probably have their own proprietary file type, right? So, this is one everybody sort of agrees on that can extract the genealogy data set right. Is that right?

Gordon: Right, and there are differences between the proprietary program and GEDCOM. There are some products out there that only support GEDCOM. So that’s their proprietary format.

Why Use a GEDCOM?

(03:45) Lisa: So why should we use one a GEDCOM. When would we find ourselves wishing we had this universal file?

Gordon: Family history is more of a record keeping whether it’s photos and stories and genealogical data. People like to keep it and have control over it. So, GEDCOM is I like the word “personal”. You can personally control it. It’s just a .GED file, so any operating system can handle copying and emailing it. So, for personal control, preservation and sharing of genealogical data. It’s the most universally accepted format.

I would think for your backup purposes because it’s so universal, make sure that the program that you’re using has the ability to save your data in GEDCOM. Then you can decide whether you put it in your thumb drive or removable drive or you put it up in the cloud, you can decide how to preserve it. Think of it more as your personal file over this important information.

(05:31) Lisa: I like that idea. I’m probably not alone in that I once had somebody give me a little floppy disk and it had the whole family tree that this person had been working on. Unfortunately, it was a proprietary file, and it was a program that no longer exists. I’m helpless to be able to use it. So, a GEDCOME can really solve that issue.

Do All Family Tree Programs Support GEDCOM?

(06:00) You kind of touched on this, but I just want to just double check. Can all family tree programs and websites export the GEDCOM? Are you familiar with anything that don’t?

Gordon: I would say all of the popular programs and websites make it possible to import GEDCOM, and most of them allow for exports. There are some exceptions to that rule. If you’re going to spend your time using a program, look to see if it’s GEDCOM compatible.

To help even more so standardize the industry, the software providers commit to implementing the newest version of GEDCOM. Much of that is backward compatible. We presented those that have or will be planning to implement the newest version of GEDCOM at Rootstech. You can search at Rootstech for “GEDCOM” and see the videos of what’s been rolled out and what’s coming.

Who Owns and Controls GEDCOM?

(07:41) Lisa: Is there one particular group or authority or somebody who’s in charge of deciding what the GEDCOM is and how it works? Or is that a role that FamilySearch is playing?

Gordon: It is a role that Family Search has been playing. FamilySearch is the software development, education marketing, support arm of the department called The Family History Department of The Church of Jesus Christ of Latter Day Saints. So sometimes because of marketing reasons, people think that we’re different. Family Search is totally run by the Church of Jesus Christ of Latter Day Saints.

From a historical standpoint, the original specification was created and released in 1984. All subsequent versions have been copyrighted by the Church of Jesus Christ of Latter Day Saints.

Now, in the last three years, as a like a product manager, I took on the responsibility for working on the new version, version 7 of GEDCOM. But it’s always been an effort of FamilySearch as the outreach arm for the Family History Department.

What we did differently in this last version is we solicited all the key players and software companies. It was much more of a collaborative effort to go through the changes, things to keep, things to just get rid of. It took about two years working with many people. Now the version is what is called a public GitHub repository. As we worked toward version 7, it was to prepare it for a starting point. The decision process is still a steering committee sponsored by FamilySearch, but the input and the communication on changes is open to all software developers. You can learn about all that because it’s hosted at GEDCOM.io. So GEDCOME.info is kind of like the general public, and GEDCOM.io is more for technical software developers.

GEDCOM Features

(10:38) Lisa: What are some of the features of GEDCOM 7? What are some of the things that you consider when you’re continuing to develop the GEDCOM?

Gordon: The process that we worked on was, I think to eliminate ambiguity, there could be different software providers that would interpret the file specifications a little bit differently. We wanted to clean up the specifications so that there would be much more, not 100%, but a much better compatibility between the people that were reading it and writing it. So, we worked very tediously on eliminating the ambiguity.

I would think that the biggest thing is, it’s become more of a storage format of photos, and records and data. Let me read something, “FamilySearch GEDCOM version 7 incorporates the added ability to include photos, and other files when users download a FamilySearch GEDCOM 7 file from a supportive family tree product.”

Your local photos can be bundled in a special file that we called GEDZIP. It’s a GEDCOM file that is a zip package. That means that anybody that unzips that package will get the GEDCOM file and all the external files associated with it and have everything be readable. It’s a packaging technique to put everything together, which really adds to this idea of a personal preservation and sharing. Now you can package everything together and preserve it and share everything that’s important to you with others.

In addition to this zipped packaging capability, notes have been expanded for more versatile use and styling of text. When you add notes, whether it’s a relationship or a location, you can actually stylize those notes now and use bold and italic.

Many tools and sample files were created to help with self-testing. It’s based upon the Apache license, which is more of a technical slant on things, but to software developers, that means it’s an open software license. There’s a public GitHub repository that you go to github.com/familysearch so that you can request and watch ongoing changes in a more of a public environment, though Family Search is still the stewards and has the final say on decisions.

So that’s what’s new. It’s more open to the public. It’s been cleaned up with some important new features.

But backward compatibility for 90% of the GEDCOMs that are out there (and the last one was 5.5.1) is still possible. But it won’t go back to 3.0, 2.0. That’s where that’s where some of the incompatibilities are, is because people are using versions that are 20 years old. And things have changed a lot in the last 25 years. We have a clean, fresh start, and a new community working on continuous improvements. But there won’t be changes because the standards shouldn’t change much. This new version 7 is going to be pretty much the same for a while everybody gets on board.

Do GEDCOMs Include Image Files of Attached Records?

(15:21) Lisa: You mentioned photographs. Would that include image files? Would that include if we downloaded an image of a genealogical record which might be a .JPEG file? Would those come along with the GEDCOM?

Gordon: Yes, absolutely. All the elements of GEDCOM have definition of how to use them. And what’s called the multimedia link, the multimedia link means you can link to local files, JPEGs, PDFs, you know, whatever they are. And if you don’t want to put it all together, you can link to files that are in the cloud, and it will remember where they are. If you package them together in a GEDZIP file, and then you unpackage it, you’ll be able to access the local image files and the local records there.

So, this idea of putting it all together, I mean, bandwidth is much better than it used to be. But still, for people that have hundreds of thousands of images. This is not the best format for that. So they can work out a strategy taking into account the cloud service they use, and which photos they will keep locally on their computer. So, they can keep track of everything, both in the cloud and on their local drive. And that can all be referenced in this new version of GEDCOM.

Is There Data Loss When Exporting a GEDCOM?

(16:59) Lisa: Excellent.

So one of the questions I’ve heard from people is that they are concerned about loss data loss. If they’re importing or exporting, maybe going back and forth, is there a chance that you’re going to lose things or even introduce an error of some type?

Gordon: This is kind of the issue of the work on version 7. One of the biggest issues is not only new features, but to get a new standard to kind of clean the slate. If you get stuff into the new GEDCOM version 7 the likelihood of data losses is greatly reduced. So, we’re encouraging the adoption and use of GEDCOM 7 because it’s less likely to cause any data loss or errors.

Family Search and industry experts have worked for two years to remove ambiguities, simplify the definitions and samples in order to eliminate the possibility of data loss and errors when transferring between programs. In the long run, not only does it include more media, but the whole goal is to improve the consistency, the compatibility and minimize or even eliminate data loss. So, what you will start being seeing is the question “is GEDCOM 7 compatible?” Because GEDCOM 7, when we were working on something that was 20 years old, is going to be more compatible in the future. We have a body to watch out for it. Your data will migrate to the new version without data loss. But looking at down the road, staying with the version 7 or higher will assure a sure better preservation of what you have.

Learn More About GEDCOM at Rootstech

(19:17) Lisa: I think you mentioned or alluded to that there were some announcements at Rootstech 2022.

Gordon: Yes, go to the sessions and type in “GEDCOM” and you will get three opportunities. One is a session called GEDCOM 7 Launched and Rolling Strong. Another session will be FamilySearch GEDCOM 7 What’s Next? And the answer is teamwork.

There’s two pre-recorded videos about the What’s New in GEDCOM 7 and then how the industry’s going to join together in working on it in the future. In in one of the sessions, the first one, there actually is a slide that shows all the companies that have committed to it. But all the majority of the companies have said, both in the cloud and desktop and laptop, and some have said when they’re going to release it. And one company I think, is announcing their release at Rootstech of the new GEDCOM version 7.

Future Updates and Changes to GEDCOM

(20:44) Lisa: That’s great to see. Anything I didn’t ask you or that you think people should really be aware of as they move forwarding and keeping up to date with GEDCOM 7?

Gordon: Again, with a standard, we don’t want to change too much too fast, because they wanted to get solid as a new transfer format.

I think the big areas that we’re working on for future versions is related quite a bit to internationalization. There are probably 20 different calendaring systems that are different than what we do in the U.S. To be able to respect those different calendars and to understand the translation between calendars is a big part of internationalizing GEDCOM.

The other part related to that is that there are some places in the world where how they define relationships between people is not typical to either the US or Western Europe. And so we are working on major upgrades and encourage people to come join with us. With naming conventions we may think given name, surname, but in reality, there’s other relationships that get into the name. If we even go to Africa their name is the first name may go back 10 generations, so their name is a memorization of all those names. So, improving on names is an important effort, the structure and relationships.

Another improvement is places. We think hierarchal and certain jurisdictions, but over time, and in different areas of the world, how you organize places is different. We need to address that in the GEDCOM specification.

Sources and Citations need to be upgraded for the genealogical community. And so, we certainly invite not only software developers, but genealogists to join our effort to improve sources and citations.

GEDCOM Hypothesis

One thing I’m really excited about is that we have a team that’s been working a year, and they’re probably working on it another year or two, on what we call hypothesis. This is so that you can share information without claiming it as a conclusion, and keep it separate from a conclusion. This encourages collaboration. So instead of arguing about I’m right, you’re wrong, we call it a hypothesis. Then we can have a discussion until there’s enough sources to prove it. This Hypothesis module I think is going to be really exciting. But that won’t be for a couple years or so until we actually release it.

Lisa: I think that’s a terrific idea because so often we are just battling with ourselves over what we think the answer is, and we want to track it while we’re doing it.

I’m curious: sometimes we go to a website, and you have to pick what language you speak. Perhaps if you’re searching for videos on YouTube you might say English. Is this something being considered? Is the goal no matter what that it’s only one type of file that serves every country or was there a consideration that you could select your country and then the GEDCOM would support your calendar and your geographic areas. I’m sure that was a discussion.

Gordon: Oh, absolutely. And, but what you’re talking about, just to be clear, is the specification to give all of the options and more to the software developer. The software developer can decide the language of the interface, and many of them are already doing this. So the actual presentation, if it’s Norwegian, or Danish, or whatever, it’s different according to the language that you place. What we’re looking according to your language of choice is that the orientations are names, relationships, and places jurisdictions, will be easy for the software developer to switch to by just changing that.

When we look at an international – how people look at information – it may be a different lens that they look through. So having the ability to give the software developers out of our future specs, to switch their interface, and switch around because they might be working in one part of the country because of their heritage, and then they might work in another and to be switched between it and to still have the data be the same, regardless of what national lens they’re looking through.

Lisa: It’s amazing that one little package contains so much and so much flexibility. That’s really terrific.

The Team Working on GEDCOM 7

(26:52) Gordon: I won’t drop names but in my immediate steering committee, that we meet with weekly, not only do I have three representations from within FamilySearch, but from the community, I like to call them doctors, they are doctors, they have their PhDs in computer science. Some are genealogists, they have their peers, one is even a linguistic professor. Another is an actual legal professional. It’s been wonderful to work with such experts, really, that are reasonable, and want to make things easy for the software developer. So, it’s quite a dilemma, instead of just making it right in the specification, but we’ve got to make it right and make it easier for the software developers to implement it. So that’s my thanks to all the people I’ve been able to work with.

GEDCOM Resources

(28:19) Lisa: Visit GEDCOM.io and GEDCOM.info.

Are they able to offer any volunteering opportunities? Do you need the help of people who are doing genealogy?

Gordon: Oh yes, you can volunteer in lots of different ways at GEDCOM.io.

Lisa: Thank you so much for taking time to explain GEDCOM.

Resources

Downloadable ad-free Show Notes handout for Premium Members. (Learn more and join Premium Membership here.)

 

What To Do If a Scrapbook Gets Wet (or Photo Album or Pictures)

water damaged scrapbookWhen family scrapbooks get wet, the result is not pretty. In fact, it can be quite dire for the scrapbook and its precious contents.

“Water can cause the bleeding of inks and dyes in journal entries, digital photographs, and decorative papers, causing them to appear blurry or streaked,” says this article in Scrapbook Retailer. “When exposed to water, some prints and materials will soften and stick to adjacent surfaces. Papers that get wet can become distorted or warped and some may even dissolve completely in water.”

Even more yucky? “Dirty water from sewage leaks, floodwaters from rivers, and colored liquids like fruit juices make the clean-up process more difficult and staining of the album materials more likely.”

Preventing the damage in the first place is of course the best option, but it’s not always an option we’re given. Floods happen. Spills happen. Windows get left open.

So what to do if a scrapbook gets wet? Or a photo album or loose pictures?

First, says the Library of Congress, “Take necessary safety precautions  if the water is contaminated with sewage or other hazards or if there is active (wet or furry) mold growth.”

“In general, wet photographs should be air dried or frozen as quickly as possible,” states the Northeast Document Conservation Center website. “Once they are stabilized by either of these methods, there is time to decide what course of action to take.” But don’t delay too long, they say. “Time is of the essence: the longer the period of time between the emergency and salvage, the greater the amount of permanent damage that will occur.”

A few more tips from that same article on the Northeast Document Conservation Center website, written by Gary Albright:

  • Save prints before plastic-based films, as the latter will last longer.
  • Allow water to drain off photos first, as needed. Then air dry photographs, face up, laying flat on paper towels. Negatives should be hung to dry.
  • Separate wet photos from each other and other items (like a scrapbook page) as much as possible.
  • If photos are stuck together, freeze them as a bunch, wrapped in wax paper. Then thaw them. As they gradually thaw, peel photos off and let them air dry.
  • Don’t worry if pictures curl up while they are drying. You can flatten them once they’re totally dry.

Unfortunately, some very old photo types will not survive a water bath at all. Others may weather a quick dip but not long-term exposure to dampness. It’s SO important to preserve images digitally! You can scan entire album pages if they fit on your scanner, so you can record captions or the arrangement of pictures on a page. Or use a scanner like Flip-Pal that has stitching software to help stitch together larger images.

In a pinch, snap pictures with your mobile device: close-ups of photographs and captions, and full-page images that at least capture how it’s laid out (even if at a lower resolution). Mobile Genealogy: How to Use Your Tablet and Smartphone for Family History Research by Lisa Louise Cooke has a chapter on digital imaging apps that can help you digitally preserve family albums and scrapbooks–whether they’ve gotten wet or not.

Christmas in July BackblazeLisa Louise Cooke trusts all our computer files–including images, sound files and videos that have taken thousands of hours to create–to Backblaze online backup service, the official backup of Genealogy Gems. For about $5 a month (or $50 for an entire year), you can protect your files, too. It only takes a couple of minutes to give yourself the peace of mind of knowing that, even if disaster strikes, you’ll still be able to recover your digital files quickly and easily. Go to www.Backblaze.com/Lisa to get started.

 

 

California in the 1940 Census

Archives.com and other community project partners recently announced the release of the California 1940 census index, and have provided a neat infographic highlighting California in 1940.

I was particularly thrilled to see the searchable California index released because my parents and all my grandparents were in the state , working farms and just a year away from going to war and building ships for the cause.

Enjoy!

1940 census archives.com

We Dig These Gems! New Genealogy Records Online

Our weekly roundup of new genealogy records online includes: the 1891 NSW Australia census;Portsmouth, England electoral registers; Frankfurt, Germany deaths; Massachusetts Revolutionary War soldiers; North Carolina probate and recent U.S. obituaries.

AUSTRALIA CENSUS. FamilySearch has added over 300k entries to its indexed records of the 1891 Australia Census for New South Wales.

ENGLAND ELECTORAL REGISTERS. Findmypast continues to expand its collection of electoral registers with nearly 200k transcripts from Portsmouth, England (1835-1873).

GERMANY DEATHS. Over half a million indexed records and accompanying images are at a new, free FamilySearch collection of death records for Frankfurt, Germany (1928-1978).

MASSACHUSETTS REVOLUTIONARY WAR. A new browsable collection of “index cards to muster rolls of soldiers who served in Massachusetts regiments during the Revolutionary War, 1775-1783″ is now searchable at FamilySearch. The card file comes from the Massachusetts State Archives in Boston.

NORTH CAROLINA PROBATE. More than a half million images and 25,000 indexed records have been added to a free collection of North Carolina estate records (1663-1979) at FamilySearch.

US OBITUARIES. FamilySearch has updated its collection of recent U.S. obituaries indexed from GenealogyBank newspaper images. Nearly 15 million records have been added. The  index is free to search.

thank you for sharingThank you for sharing these new genealogy records online with your fellow genealogy buddies and society members! You’re a gem!

Knowles Jewish Genealogy Collection Has Over a Million Entries

Ketubah Circa 1860. This is the ketubah (marriage contract) of Hannah and Hayyim from their marriage on Tuesday, April 6, 1886 (א׳ ניסן תרמ״ו) in the town of Brody. Image by Yoel Ben-Avraham on Flickr Creative Commons at https://www.flickr.com/photos/epublicist/1355967207/in/photolist-.

Ketubah Circa 1860.
This is the ketubah (marriage contract) of Hannah and Hayyim from their marriage on Tuesday, April 6, 1886 (א׳ ניסן תרמ״ו) in the town of Brody. Image by Yoel Ben-Avraham on Flickr Creative Commons at https://www.flickr.com/photos/epublicist/1355967207/in/photolist-.

Looking for an online resource of Jewish family trees?

“The Knowles Collection, a quickly growing, free online Jewish genealogy database linking generations of Jewish families from all over the world, reached its one-millionth record milestone and is now easily searchable online,” says a recent FamilySearch press release.

“The collection started from scratch just over seven years ago, with historical records gathered from FamilySearch’s collections. Now the vast majority of new contributions are coming from families and private archives worldwide. The free collection can be accessed at FamilySearch.org/family-trees.

According to FamilySearch, “The databases from the Knowles Collection are unlike other collections in that people are linked as families and the collection can be searched by name, giving researchers access to records of entire families. All records are sourced and show the people who donated the records so cousins can contact one another. New records are added continually, and the collection is growing by about 10,000 names per month from over 80 countries. Corrections are made as the need is found, and new links are added continually.”

The database was started by Todd Knowles, a Jewish genealogy expert at the Family History Library in Salt Lake City. Jewish communities from around the world have added to it: “The Knowles Collection has grown from Jews of the British Isles (now with 208,349 records), to Jews of North America (489,400), Jews of Europe (380,637), Jews of South America and the Caribbean (21,351), Jews of Africa, the Orient, and the Middle East (37,618), and the newest one, Jews of the Southern Pacific (21,518).” Keep up with the Knowles Jewish Collection at its blog.

Pin It on Pinterest

MENU