## Abstract We summarize the work done by the contributors to Group 13 at Genetic Analysis Workshop 17 (GAW17) and provide a synthesis of their data analyses. The Group 13 contributors used a variety of approaches to test associations of both rare variants and common singleโnucleotide polymorphisms
A dictionary model for haplotyping, genotype calling, and association testing
โ Scribed by Kristin L. Ayers; Chiara Sabatti; Kenneth Lange
- Publisher
- John Wiley and Sons
- Year
- 2007
- Tongue
- English
- Weight
- 182 KB
- Volume
- 31
- Category
- Article
- ISSN
- 0741-0395
No coin nor oath required. For personal study only.
โฆ Synopsis
Abstract
We propose a new method for haplotyping, genotype calling, and association testing based on a dictionary model for haplotypes. In this framework, a haplotype arises as a concatenation of conserved haplotype segments, drawn from a predefined dictionary according to segment specific probabilities. The observed data consist of unphased multimarker genotypes gathered on a random sample of unrelated individuals. These genotypes are subject to mutation, genotyping errors, and missing data. The true pair of haplotypes corresponding to a person's multimarker genotype is reconstructed using a Markov chain that visits haplotype pairs according to their posterior probabilities. Our implementation of the chain alternates Gibbs steps, which rearrange the phase of a single marker, and Metropolis steps, which swap maternal and paternal haplotypes from a given maker onward. Output of the chain include the most likely haplotype pairs, the most likely genotypes at each marker, and the expected number of occurrences of each haplotype segment. Reconstruction accuracy is comparable to that achieved by the best existing algorithms. More importantly, the dictionary model yields expected counts of conserved haplotype segments. These imputed counts can serve as genetic predictors in association studies, as we illustrate by examples on cystic fibrosis, Friedreich's ataxia, and angiotensinโI converting enzyme levels. Genet. Epidemiol. ยฉ 2007 WileyโLiss, Inc.
๐ SIMILAR VOLUMES
## Abstract In caseโcontrol single nucleotide polymorphism (SNP) data, the allele frequency, Hardy Weinberg Disequilibrium, and linkage disequilibrium (LD) contrast tests are three distinct sources of information about genetic association. While all three tests are typically developed in a retrospe
## Abstract Estimation and testing of genetic effects (genotype relative risks) are often performed conditionally on parental genotypes, using data from caseโparent trios. This strategy avoids having to estimate nuisance parameters such as parental mating type frequencies, and also avoids generatin