HGDP

Posted by Zack on January 25, 2011

Human Genome Diversity Project (HGDP) is the best resource for a diverse set of genomic data. It has 1050 individuals from 52 different populations.

I got the Stanford University data which has data for 660,918 SNPs from 1,043 samples. It is claimed that the forward strand is given but that turned out not to be true and I had to flip strands and make sure I didn't include any ambiguous A/T or C/G strands in my dataset.

I followed the recommendations of Rosenberg (spreadsheet) in excluding some atypical samples and relatives, leaving me with 940 samples.

I also excluded the Native American samples because we are not interested in them and they are very closely related either due to recent endogamy or ancient bottlenecks. (yeah I had the nerve to write that.)

Of the total of 876 samples, here are the numbers for our populations of interest:

Total South Asians	190
Balochi	24
Brahui	25
Burusho	25
Hazara	22
Kalash	23
Makrani	25
Pathan	22
Sindhi	24

These samples have about 541,560 SNPs in common with 23andme v2.

Datasetbaloch, brahui, burusho, data, genome, hazara, hgdp, kalash, makrani, pathan, sindhi

← 23andme v3 Data

SGVP →

3 Comments.

Admixture: Reference Population | Harappa Ancestry Project - pingback on January 29, 2011 at 9:15 am
HGDP to PED Conversion | Harappa Ancestry Project - pingback on February 5, 2011 at 6:59 pm
My Biogeographical Ancestry | Procrastination - pingback on February 11, 2011 at 6:20 am

Trackbacks and Pingbacks:

Admixture: Reference Population | Harappa Ancestry Project - Pingback on 2011/01/29/ 09:15
HGDP to PED Conversion | Harappa Ancestry Project - Pingback on 2011/02/05/ 18:59
My Biogeographical Ancestry | Procrastination - Pingback on 2011/02/11/ 06:20

Harappa Ancestry Project

Genetics and South Asia

HGDP

Related

3 Comments.

Trackbacks and Pingbacks:

Contact

My Sites

Data

Affiliate DNA Tests

Categories

Archives

Recent Comments

Blogroll

Harappa Ancestry Project

Genetics and South Asia

HGDP

Share this:

Related

3 Comments.

Trackbacks and Pingbacks:

Contact

My Sites

Data

Affiliate DNA Tests

Categories

Tags

Archives

Recent Comments

Blogroll