A dynamic programming algorithm for haplotype block partitioning
Top Cited Papers
Open Access
- 21 May 2002
- journal article
- Published by Proceedings of the National Academy of Sciences in Proceedings of the National Academy of Sciences
- Vol. 99 (11) , 7335-7339
- https://doi.org/10.1073/pnas.102186799
Abstract
We develop a dynamic programming algorithm for haplotype block partitioning to minimize the number of representative single nucleotide polymorphisms (SNPs) required to account for most of the common haplotypes in each block. Any measure of haplotype quality can be used in the algorithm and of course the measure should depend on the specific application. The dynamic programming algorithm is applied to analyze the chromosome 21 haplotype data of Patil et al. [Patil, N., Berno, A. J., Hinds, D. A., Barrett, W. A., Doshi, J. M., Hacker, C. R., Kautzer, C. R., Lee, D. H., Marjoribanks, C., McDonough, D. P., et al. (2001) Science 294, 1719–1723], who searched for blocks of limited haplotype diversity. Using the same criteria as in Patil et al., we identify a total of 3,582 representative SNPs and 2,575 blocks that are 21.5% and 37.7% smaller, respectively, than those identified using a greedy algorithm of Patil et al. We also apply the dynamic programming algorithm to the same data set based on haplotype diversity. A total of 3,982 representative SNPs and 1,884 blocks are identified to account for 95% of the haplotype diversity in each block.Keywords
This publication has 11 references indexed in Scilit:
- Bayesian Haplotype Inference for Multiple Linked Single-Nucleotide PolymorphismsAmerican Journal of Human Genetics, 2002
- Blocks of Limited Haplotype Diversity Revealed by High-Resolution Scanning of Human Chromosome 21Science, 2001
- Genetic variation in the 5q31 cytokine gene cluster confers susceptibility to Crohn diseaseNature Genetics, 2001
- Linkage disequilibrium in the human genomeNature, 2001
- A New Statistical Method for Haplotype Reconstruction from Population DataAmerican Journal of Human Genetics, 2001
- Variation is the spice of lifeNature Genetics, 2001
- Extent and Distribution of Linkage Disequilibrium in Three Genomic RegionsAmerican Journal of Human Genetics, 2001
- Haplotype Structure and Population Genetic Inferences from Nucleotide-Sequence Variation in Human Lipoprotein LipaseAmerican Journal of Human Genetics, 1998
- Maximum-likelihood estimation of molecular haplotype frequencies in a diploid population.Molecular Biology and Evolution, 1995
- HAPLO: A Program Using the EM Algorithm to Estimate the Frequencies of Multi-site HaplotypesJournal of Heredity, 1995