Unknown

Dataset Information

0

Adjusting family relatedness in data-driven burden test of rare variants.


ABSTRACT: Family data represent a rich resource for detecting association between rare variants (RVs) and human traits. However, most RV association analysis methods developed in recent years are data-driven burden tests which can adaptively learn weights from data but require permutation to evaluate significance, thus are not readily applicable to family data, because random permutation will destroy family structure. Direct application of these methods to family data may result in a significant inflation of false positives. To overcome this issue, we have developed a generalized, weighted sum mixed model (WSMM), and corresponding computational techniques that can incorporate family information into data-driven burden tests, and allow adaptive and efficient permutation test in family data. Using simulated and real datasets, we demonstrate that the WSMM method can be used to appropriately adjust for genetic relatedness among family members and has a good control for the inflation of false positives. We compare WSMM with a nondata-driven, family-based Sequence Kernel Association Test (famSKAT), showing that WSMM has significantly higher power in some cases. WSMM provides a generalized, flexible framework for adapting different data-driven burden tests to analyze data with any family structures, and it can be extended to binary and time-to-onset traits, with or without covariates.

SUBMITTER: Zhang Q 

PROVIDER: S-EPMC4236253 | biostudies-literature | 2014 Dec

REPOSITORIES: biostudies-literature

altmetric image

Publications

Adjusting family relatedness in data-driven burden test of rare variants.

Zhang Qunyuan Q   Wang Lihua L   Koboldt Dan D   Boreki Ingrid B IB   Province Michael A MA  

Genetic epidemiology 20140828 8


Family data represent a rich resource for detecting association between rare variants (RVs) and human traits. However, most RV association analysis methods developed in recent years are data-driven burden tests which can adaptively learn weights from data but require permutation to evaluate significance, thus are not readily applicable to family data, because random permutation will destroy family structure. Direct application of these methods to family data may result in a significant inflation  ...[more]

Similar Datasets

| S-EPMC4143885 | biostudies-literature
| S-EPMC4143729 | biostudies-literature
| S-EPMC3201701 | biostudies-literature
| S-EPMC2811752 | biostudies-literature
| S-EPMC4940283 | biostudies-literature
| S-EPMC6174288 | biostudies-literature
| S-EPMC2912645 | biostudies-literature
| S-EPMC4649945 | biostudies-literature
| S-EPMC4454531 | biostudies-literature