Showing posts with label resources. Show all posts
Showing posts with label resources. Show all posts

Thursday, June 25, 2020

DNA Case Study: Lemuel Patchen and Limits of Autosomal DNA Testing

[Minor corrections made 31 August 2020.]

We've traced our Patchen ancestors back to a Lemuel Patchen in Ontario, Canada in 1820, and his son, Thomas. Thomas was born in Canada in about 1796. Other than one census record, there is no other information on this pair in Canada. Other Patchen researchers speculated that our Lemuel was the same who had abandoned his family in the early 1790s and headed into Canada. This Lemuel was part of the extensively researched Patchen family of Connecticut. I described the details several years ago in another blog post: http://ourfamilyforest.blogspot.com/2012/07/lemuel-patchen-1770-1850s.html .

Recently, I had my DNA analyzed at Ancestry.com, and have found several DNA matches to Patchen descendants. Four of them are descendants of Thomas and, since I have quite a bit of information on our Patchens, were easily placed in my family tree as 3rd and 4th cousins. Four others, though, seem to have genealogies connecting them to the Patchens of Connecticut. Two are descendants of Walter Lockwood Patchen, a brother of Lemuel Patchen, both sons of George Patchen, born in 1737 in Wilton, Connecticut.  If we are descended from this Lemuel, these DNA matches are 6th cousins of mine. Two others are descendants of Ann Patchen Morehouse who, according to the extensively researched genealogy, is the daughter of Jabez Patchen, a first cousin to George. This would make me an 8th cousin to these DNA matches. I will mention, though, that there are some who argue that Ann Morehouse was not the daughter of Jabez, that her father was actually George, father of Lemuel and Walter. So her descendants may actually be 6th cousins, also.

Can the DNA analysis tell me if our Lemuel is the son of George Patchen of the Connecticut Patchens?

To answer that, lets look at the numbers. All four of the Connecticut Patchens share 8cM of DNA with me, and all are estimated to be somewhere between  5th and 8th cousins. The good new is that 8cM (cM indicate how likely it is that DNA is inherited), while small, is not insignificant. So it is likely, especially with several matches, that we are related to the Connecticut Patchens. Is our Lemuel the son of George Patchen, who left to Canada? There are some useful charts that might help.

A good resource for using DNA for genealogy is the International Society of Genetic Genealogy (ISOGG). A table on their statistics wiki page (Average autosomal DNA shared by pairs of relatives) shows how many cM of DNA are expected to be shared for different relationships. There is a lot of variability in the amount of DNA inherited from a specific ancestor, so the numbers in this table are the expected average values. The last line of this table shows that 3.32cM, on average, will be shared by 5th cousins. If you read the whole page, or study the table, you'll see that the average is divided in half for each additional generation of ancestor. For example, 4th cousins share 1/4 as much DNA as 3rd cousins. In the table 3rd cousins share 53cM of DNA, on average, and 4th cousins share about 13cM of DNA, about 1/4 as much. In this Patchen example, we're looking at 6th or 8th cousins, so take the last line of the table (5th cousins share on average 3.32cM of DNA) and divide repeatedly by four to see that sixth cousins share about 0.8cM, seventh cousins about 0.2cM, and eighth cousins about 0.05cM. Compare this to the measured 8cM DNA shared by me and my Connecticut Patchen matches. We share at least 10x more DNA than expected for the 6th or 8th cousin relationship I was considering. This implies we are much more closely related, but I know from our family trees (assuming they are accurate) that we are not more closely related.

When you study distant relationships, say more distant than 4th cousin (expected 13cM shared DNA), we run into a problem. Very small amounts of DNA may be the same between individuals, but not because it is inherited. They may be randomly the same. Some may be related to communities in which individuals lived. There may be errors in detecting. Or other reasons that I don't know about. But because very small segments that match may not be inherited from individual ancestors, and we can't know which are inherited and which are not, testing companies use a threshold when reporting shared DNA, usually 6 to 8cM. Because of this, when comparing distant relatives, many of the small dna segments are removed because they are below the threshold. In this case of 6th and 8th cousins, whose expected shared DNA is 0.8cM and 0.2cM, both well below the rejection threshold, we only see those relatives who are sharing much more than the average expected. There are two effects of this. First, most of the distant matches are below threshold so aren't even shown as matches. Second, those that do exceed the threshold are only those that share significantly more than the average, so the shared DNA will seem high. My Connecticut Patchen matches should share less than 1cM, but are measured as 8cM. So still can't tell what my relationship is to these Patchen matches. (But I'm pretty sure they are relatives.)

ISOGG's Cousin Statistics table shows the first effect. You can see that Ancestry can only detect about 11%  (about 1/10) of 6th cousins, and less than 1% (1 in 100!) of 8th cousins. The second effect is shown by the Shared cM Project of Dr. Blaine Bettinger, summarized in the table below. He gathers data from people who have DNA analyzed about their known relationships to DNA matches and the amount of DNA shared. The recent 2020 update summarizes over 60,000 data submissions. He generates a report that contains lots of useful charts, but the main one is this (click on it to make it bigger):


This chart shows what the actual reported amounts of DNA are for various relationship. So, for example, the ISOGG first chart shows that first cousins share on average 850cM of DNA. Bettinger's chart shows us that for first cousins (green box labeled 1C next to the central SELF box) companies are actually measuring an average of 866cM. But look at 5th cousins. ISOGG/theory tells us to expect about 3.3cM shared DNA. Bettinger reports that 5th cousins are reported, on average, as sharing 25cM, about 8 times what is expected. This is probably in large part because if the average is 3.3cM, but there is lots of variability above and below this, and everything below about 7cM (twice the expected average) is not considered, the reported average number will be much higher than the expected average. It is also likely for distant relationships that there are multiple relationships, each contributing some DNA, some of which the DNA matches don't know about.

So does this table help determine my Patchen relationships? According to this Shared cM Project table, 6th and 8th cousins are reporting, on average, 18cM and 11cM shared DNA, respectively. This chart says it's more likely that my Connecticut Patchen matches, all of which share 8cM with me, are 8th cousins. But if our Lemuels are the same person, which I think is true, two of these matches are known to be 6th cousins. How can that be? Take another look at the above chart. For 6th cousins, the range reported was 0 (in other words, not detected as a match at all) to 71cM shared. For 8th cousins, it was 0 to 42cM shared. So my 8cM matches could be in either one of these ranges. There are other numbers from the Shared cM Project (standard deviations) that I can use to nudge my opinion about these relationships, but while I can be confident that we are related, I can't identify the exact relationship.

That's a lot of work and explanation for a shoulder shrug, but it demonstrates limitations of autosomal DNA testing, especially for distant cousins, it showed how some useful tables and charts can be used in testing a relationship hypothesis, and it does show some evidence that our Lemuels are the same.

Sunday, September 1, 2019

Caseys in Galbally, Limerick, Ireland; Research resources

I recently made a DNA connection to a Casey family, and now am fairly certain that what I had suspected from census records, that Patrick Casey (b. ca 1801 in Ireland, married Hanora Norris in Galbally, where most/all of their kids were baptized) was a brother of Catherine Casey Cussen. This is just a note about some resources.

I've spent a lot of time going through church register images for the parish of Galbally, on the National Library of Ireland site ( https://registers.nli.ie/parishes/0264 ). The images are not indexed, so searching is like what we used to do when searching through census and newspaper films at local libraries. Except I can do this on my computer at home. I thought this would be a fairly quick job, but it turns out to be enormous. I'm looking for all Cushing and Casey entries to get a pool of candidates for the family in Ireland. It turns out there are about 500 images, most containing two pages from a register. I'm finding about two or three of interest per image. A baptismal record is typically a date, the child, two parents, two godparents, a page number, sometimes a note about the father's profession or town of residence, or that the child was "illegitimate", so typically about nine fields of information, often difficult to read. A marriage record is the married couple, two witnesses, a date (three fields), with occasional notes and a page number. At the end I add a film numbers, too, so that I can easily find the record again, so the whole is typically eight fields of information. That comes out to an estimate of about of about 2500 records and 20,000 recorded fields of information. So I should have expected a lot of work. I think I'm about halfway done.

The interesting part of this near drudgery is seeing all the names, something of a directory of neighbors of my Casey & Cushing ancestors. Many of the names are familiar as spouses of marriages that took place after immigration to the US, so I wonder if many of the Cushing & Casey kids and grandkids married into families the parents knew from "the old country". I've also seen some of these names in DNA matches to my dad, which opens some paths of searching for common ancestors. Some of the names that were very common in the Galbally register were Barry, Blackburn, Bourke, Brien, Butler, Byrrane, Carty, Casey, Clancy, Condon, Connor, Cronin, Cummings, Cunningham, Cussen/Cushen/Quishian, Dalton, Dawson, Dea, Donohoe, Dunn, Dwyer, Fitzgerald, Fogarty, Fraher, Fruin, Gorman, Grafton, Halloway, Hanrahan, Hayes, Heffernan, Henebry, Hennesy, Ivory, Kiely, Kirby, Landers, Lynch, Mahoney, Mara, Martin, Megrath, Moloney, Mullins, Murphy, Neil, Noonan, Picket, Power, Quain, Ryan, Sampson, Sheehan, Slattery, Sullivan, Walsh. And many of these added an O (O'Brien, O'Neal, O'Sullivan ...) or a Mc (McCarthy, McGrath, ...).

Another site I found interesting is the Irish Placenames Database at https://www.logainm.ie/en/s?txt=galbally&str=on . My browser identifies this as Dublin City University, but I don't know what exactly the project is. Often a register record would have a place name associated with a groom or a father, and the strange name and difficult-to-read writing made it difficult to record a meaningful place name. I didn't have a lot of success, but I found the resource interesting for locating on a map Irish place names more generally. This seems to be related to a project to preserve Irish culture by identifying and officially recognized places.

At the top of web page are links to what seem to be (a brief glance) other Irish collections. Above and to the left of the map is a link to "Meitheal Logainm.ie", which seems to be a place for people to submit local place names that may not be officially recognized, yet. But it's also searchable. I don't see any descriptions, but there are lots of places identified if you zoom in close. Some of the site are in the Irish language. ainm.ie seems to be a collection of biographies, but only in Irish. https://www.duchas.ie/en/ is a site collecting items to preserve Irish culture, through stories and photos. For example, I found this in their schools collection: https://www.duchas.ie/en/cbes/4922055/4848074/5009531 giving a local explanation of Galbally, which apparently means "town of the strangers".

A last resource, not new but perhaps you haven't seen it, is built around Matheson's statistics (published in a book that people have found very useful) about the Surnames of Ireland. I don't want to go look up the book right now, but from memory he summarized an enumeration of the births that took place in about 1890 throughout Ireland, and it is widely used to find to find families to help focus genealogy research to more likely areas of the country. The country had been decimated by famine related emigration, so the numbers and distributions of names aren't the same as they were in the 1830s and pre-famine 1840s, when most of my Irish ancestors lived there, but it is a valuable resource. Many of us bought the book to look through the tables of names, but now it is searchable online at https://www.ancestryireland.com/family-records/distribution-of-surnames-in-ireland-1890-mathesons-special-report/ . I had to try several spelling variants for Cushen to find the entry in their table, so you might not find your name on a first try. The book will show you all the variants that made up the head count. The book also has some explanation of origins of some names.

I'll post other resources as I come across them. Please post your own in the comments.

Enjoy!