Display Settings:

Format

Send to:

Choose Destination

    Am J Hum Genet. 2007 Nov;81(5):1084-97. Epub 2007 Sep 21.

    Rapid and accurate haplotype phasing and missing-data inference for whole-genome association studies by use of localized haplotype clustering.

    Browning SR, Browning BL.

    Department of Statistics, The University of Auckland, Auckland, New Zealand. s.browning@auckland.ac.nz

    Whole-genome association studies present many new statistical and computational challenges due to the large quantity of data obtained. One of these challenges is haplotype inference; methods for haplotype inference designed for small data sets from candidate-gene studies do not scale well to the large number of individuals genotyped in whole-genome association studies. We present a new method and software for inference of haplotype phase and missing data that can accurately phase data from whole-genome association studies, and we present the first comparison of haplotype-inference methods for real and simulated data sets with thousands of genotyped individuals. We find that our method outperforms existing methods in terms of both speed and accuracy for large data sets with thousands of individuals and densely spaced genetic markers, and we use our method to phase a real data set of 3,002 individuals genotyped for 490,032 markers in 3.1 days of computing time, with 99% of masked alleles imputed correctly. Our method is implemented in the Beagle software package, which is freely available.

    PMID: 17924348 [PubMed - indexed for MEDLINE]

    PMCID: 2265661

    Supplemental Content

    Click here to read Click here to read Click here to read Click here to read