Showing posts with label FTDNA. Show all posts
Showing posts with label FTDNA. Show all posts

Thursday, March 24, 2011

FTDNA Family Finder Illumina data for project members

If:
  • you have already submitted your data to the Project have received your id (FFD)
  • you have not submitted 23andMe data
  • you have received your new Illumina data from FTDNA
then, you can send me your new Illumina data. This will allow me to incorporate you together with the 23andMe submitters and I will compute new admixture proportions for you. Moreover, in the future you will be included in analyses that were previously reserved for 23andMe submitters due to the different chip technology.

Please reference your FFD when you submit your data.

Tuesday, January 11, 2011

Results for FFD056 to FFD061 posted

Note that I am not accepting Family Finder data at this time; these samples have accumulated since the last submission opportunity. Feel free to follow the blog to be alerted for new submission opportunities, and if you received your results, please take the time to leave a comment in the ancestry thread.

Admixture proportions can be found in the spreadsheet

All populations:

Individual bars:

Monday, December 6, 2010

Clusters galore, K=50 for Dodecad Project members (up to FFD055)

I am now repeating the Clusters galore analysis for Family Finder data (for more info, see the previous post on 23andMe data).

With 14 MDS dimensions retained, there were 50 clusters inferred in the optimal solution by MCLUST.

The results spreadsheet has rows for the 54 project participants in the first rows: each row is the probability that you belong to a particular cluster. This is followed by the reference populations where each row has the number of individuals (for that populations) that is assigned to a particular cluster.

There are also some outliers in this analysis:
FFD002 FFD004 FFD007 FFD012 FFD015 FFD016 FFD021 FFD022 FFD023 FFD038 FFD046
Check what an outlier is in the context of this analysis, and what it means.

Interestingly, because of the smaller number of Family Finder participants some previously defined clusters (for 23andMe data) such as the "Finnish" cluster do not appear here. This is not surprising at all, because for a cluster to be defined several individuals from that population must be present in the data.

Many continental Europeans of this type ended up in cluster #2. Some others, like FFD048 who is Lithuanian were assigned to the proper cluster #9, centered on Lithuanians.

This underscores the importance of having more people join the Project at the next available opportunity. This will not only create new clusters for individuals who are currently the only representatives of their populations, but it may also split already existing clusters if regional sub-populations are detected.

It is also important for project participants to drop a note at the ancestry thread, to help others make better sense of their results.

Sunday, November 28, 2010

Submission of Family Finder data is now CLOSED

Thank you all for submitting your data; the remaining results will be posted in the blog over the next week or so.

If you have submitted your data in time, but did not receive an ID yet, you will.

If you want to be alerted for future opportunities, and to keep up with the progress of the Project, please subscribe to the feed.

Thursday, November 25, 2010

ADMIXTURE analysis for Family Finder (FTDNA) samples

I am now able to provide K=10 analysis to Family Finder customers.

The rules of participation are:
  • No relatives (up to 2nd cousin)
  • 100% Eurasian or North African ancestry; the test does not include Native American samples, or an assessment of Native American ancestry
  • Data must be received by Sunday 29/11
In order to participate, you must send to dodecad@gmail.com your autosomal data (.gz ending) that you can download from FTDNA, as well as information about your known ancestry (such as country of origin, or ethnic affiliation)

There may be other opportunities for people to participate, so please subscribe to the feed.

Your raw data or genealogical information will not be shared or distributed in any manner, and it will not be analyzed for any other purpose than assessment of ancestry (i.e., not for any physical or health-related traits). It will be identified by a unique ID, known to you and me, and results will be posted in the blog using that ID. I will continue to analyze your data for ancestry, and new results will be posted using that same ID. Also, I will report aggregate results for populations with at least 5 participants.

You will receive your ancestral proportions from 10 inferred ancestral components as in the following figure:


This was generated using the same 104,790 markers that I will be using to analyze your sample. Exact admixture proportions for these populations can be found in the population spreadsheet.

Note that these proportions are not directly comparable with those using 23andMe data, as a different set of markers is used in the latter, and there is a smaller overlap between Family Finder data and those of the reference populations I am using. Here is the current population spreadsheet for 23andMe data.

There are already two Family Finder participants in the Project, with IDs FFD001 and FFD002; these are volunteers who helped me with their data when I adapted EURO-DNA-CALC for Family Finder data. Their results are in the individual spreadsheet.