Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Fast and accurate imputation of summary statistics enhances evidence of functional enrichment.

Literature DB >> 24990607

Fast and accurate imputation of summary statistics enhances evidence of functional enrichment.

Bogdan Pasaniuc¹, Noah Zaitlen², Huwenbo Shi², Gaurav Bhatia³, Alexander Gusev³, Joseph Pickrell¹, Joel Hirschhorn², David P Strachan², Nick Patterson², Alkes L Price³.

Abstract

MOTIVATION: Imputation using external reference panels (e.g. 1000 Genomes) is a widely used approach for increasing power in genome-wide association studies and meta-analysis. Existing hidden Markov models (HMM)-based imputation approaches require individual-level genotypes. Here, we develop a new method for Gaussian imputation from summary association statistics, a type of data that is becoming widely available.
RESULTS: In simulations using 1000 Genomes (1000G) data, this method recovers 84% (54%) of the effective sample size for common (>5%) and low-frequency (1-5%) variants [increasing to 87% (60%) when summary linkage disequilibrium information is available from target samples] versus the gold standard of 89% (67%) for HMM-based imputation, which cannot be applied to summary statistics. Our approach accounts for the limited sample size of the reference panel, a crucial step to eliminate false-positive associations, and it is computationally very fast. As an empirical demonstration, we apply our method to seven case-control phenotypes from the Wellcome Trust Case Control Consortium (WTCCC) data and a study of height in the British 1958 birth cohort (1958BC). Gaussian imputation from summary statistics recovers 95% (105%) of the effective sample size (as quantified by the ratio of [Formula: see text] association statistics) compared with HMM-based imputation from individual-level genotypes at the 227 (176) published single nucleotide polymorphisms (SNPs) in the WTCCC (1958BC height) data. In addition, for publicly available summary statistics from large meta-analyses of four lipid traits, we publicly release imputed summary statistics at 1000G SNPs, which could not have been obtained using previously published methods, and demonstrate their accuracy by masking subsets of the data. We show that 1000G imputation using our approach increases the magnitude and statistical evidence of enrichment at genic versus non-genic loci for these traits, as compared with an analysis without 1000G imputation. Thus, imputation of summary statistics will be a valuable tool in future functional enrichment analyses.
AVAILABILITY AND IMPLEMENTATION: Publicly available software package available at http://bogdan.bioinformatics.ucla.edu/software/. CONTACT: bpasaniuc@mednet.ucla.edu or aprice@hsph.harvard.edu SUPPLEMENTARY INFORMATION: Supplementary materials are available at Bioinformatics online.

Entities: Chemical

Mesh：

Year: 2014 PMID： 24990607 PMCID： PMC4184260 DOI： 10.1093/bioinformatics/btu416

Source DB: PubMed Journal: Bioinformatics ISSN： 1367-4803 Impact factor: 6.937

39 in total

1. Asking for more.

Authors:
Journal: Nat Genet Date: 2012-06-27 Impact factor: 38.330

Review 2. Genotype imputation for genome-wide association studies.

Authors: Jonathan Marchini; Bryan Howie
Journal: Nat Rev Genet Date: 2010-07 Impact factor: 53.242

3. Low-coverage sequencing: implications for design of complex trait association studies.

Authors: Yun Li; Carlo Sidore; Hyun Min Kang; Michael Boehnke; Gonçalo R Abecasis
Journal: Genome Res Date: 2011-04-01 Impact factor: 9.043

4. Phasing of many thousands of genotyped samples.

Authors: Amy L Williams; Nick Patterson; Joseph Glessner; Hakon Hakonarson; David Reich
Journal: Am J Hum Genet Date: 2012-08-10 Impact factor: 11.025

5. MaCH: using sequence and genotype data to estimate haplotypes and unobserved genotypes.

Authors: Yun Li; Cristen J Willer; Jun Ding; Paul Scheet; Gonçalo R Abecasis
Journal: Genet Epidemiol Date: 2010-12 Impact factor: 2.135

6. ProbABEL package for genome-wide association analysis of imputed data.

Authors: Yurii S Aulchenko; Maksim V Struchalin; Cornelia M van Duijn
Journal: BMC Bioinformatics Date: 2010-03-16 Impact factor: 3.169

7. Fast and accurate genotype imputation in genome-wide association studies through pre-phasing.

Authors: Bryan Howie; Christian Fuchsberger; Matthew Stephens; Jonathan Marchini; Gonçalo R Abecasis
Journal: Nat Genet Date: 2012-07-22 Impact factor: 38.330

8. DIST: direct imputation of summary statistics for unmeasured SNPs.

Authors: Donghyung Lee; T Bernard Bigdeli; Brien P Riley; Ayman H Fanous; Silviu-Alin Bacanu
Journal: Bioinformatics Date: 2013-08-28 Impact factor: 6.937

9. An integrated map of genetic variation from 1,092 human genomes.

Authors: Goncalo R Abecasis; Adam Auton; Lisa D Brooks; Mark A DePristo; Richard M Durbin; Robert E Handsaker; Hyun Min Kang; Gabor T Marth; Gil A McVean
Journal: Nature Date: 2012-11-01 Impact factor: 49.962

10. All SNPs are not created equal: genome-wide association studies reveal a consistent pattern of enrichment among functionally annotated SNPs.

Authors: Andrew J Schork; Wesley K Thompson; Phillip Pham; Ali Torkamani; J Cooper Roddey; Patrick F Sullivan; John R Kelsoe; Michael C O'Donovan; Helena Furberg; Nicholas J Schork; Ole A Andreassen; Anders M Dale
Journal: PLoS Genet Date: 2013-04-25 Impact factor: 5.917

86 in total

1. A simple and accurate method to determine genomewide significance for association tests in sequencing studies.

Authors: Dan-Yu Lin
Journal: Genet Epidemiol Date: 2019-01-08 Impact factor: 2.135

2. DISSCO: direct imputation of summary statistics allowing covariates.

Authors: Zheng Xu; Qing Duan; Song Yan; Wei Chen; Mingyao Li; Ethan Lange; Yun Li
Journal: Bioinformatics Date: 2015-03-24 Impact factor: 6.937

3. Partitioning heritability of regulatory and cell-type-specific variants across 11 common diseases.

Authors: Alexander Gusev; S Hong Lee; Gosia Trynka; Hilary Finucane; Bjarni J Vilhjálmsson; Han Xu; Chongzhi Zang; Stephan Ripke; Brendan Bulik-Sullivan; Eli Stahl; Anna K Kähler; Christina M Hultman; Shaun M Purcell; Steven A McCarroll; Mark Daly; Bogdan Pasaniuc; Patrick F Sullivan; Benjamin M Neale; Naomi R Wray; Soumya Raychaudhuri; Alkes L Price
Journal: Am J Hum Genet Date: 2014-11-06 Impact factor: 11.025

4. RAISS: robust and accurate imputation from summary statistics.

Authors: Hanna Julienne; Huwenbo Shi; Bogdan Pasaniuc; Hugues Aschard
Journal: Bioinformatics Date: 2019-11-01 Impact factor: 6.937

5. A powerful subset-based method identifies gene set associations and improves interpretation in UK Biobank.

Authors: Diptavo Dutta; Peter VandeHaar; Lars G Fritsche; Sebastian Zöllner; Michael Boehnke; Laura J Scott; Seunggeun Lee
Journal: Am J Hum Genet Date: 2021-03-16 Impact factor: 11.025

6. Imaging-wide association study: Integrating imaging endophenotypes in GWAS.

Authors: Zhiyuan Xu; Chong Wu; Wei Pan
Journal: Neuroimage Date: 2017-07-20 Impact factor: 6.556

7. Widespread Allelic Heterogeneity in Complex Traits.

Authors: Farhad Hormozdiari; Anthony Zhu; Gleb Kichaev; Chelsea J-T Ju; Ayellet V Segrè; Jong Wha J Joo; Hyejung Won; Sriram Sankararaman; Bogdan Pasaniuc; Sagiv Shifman; Eleazar Eskin
Journal: Am J Hum Genet Date: 2017-05-04 Impact factor: 11.025

8. Significance Testing for Allelic Heterogeneity.

Authors: Yangqing Deng; Wei Pan
Journal: Genetics Date: 2018-06-29 Impact factor: 4.562

9. Prioritizing Crohn's disease genes by integrating association signals with gene expression implicates monocyte subsets.

Authors: Kyle Gettler; Mamta Giri; Ephraim Kenigsberg; Jerome Martin; Ling-Shiang Chuang; Nai-Yun Hsu; Lee A Denson; Jeffrey S Hyams; Anne Griffiths; Joshua D Noe; Wallace V Crandall; David R Mack; Richard Kellermayer; Clara Abraham; Gabriel Hoffman; Subra Kugathasan; Judy H Cho
Journal: Genes Immun Date: 2019-01-29 Impact factor: 2.676

10. Improving Imputation Accuracy by Inferring Causal Variants in Genetic Studies.

Authors: Yue Wu; Farhad Hormozdiari; Jong Wha J Joo; Eleazar Eskin
Journal: J Comput Biol Date: 2018-10-01 Impact factor: 1.479