A Statistical Framework for Cross-Tissue Transcriptome-Wide Association Analysis.
Yiming Hu,Mo Li,Qiongshi Lu,Haoyi Weng,Jiawei Wang,Seyedeh M Zekavat,Zhaolong Yu,Boyang Li,Jianlei Gu,Sydney Muchnik,Yu Shi,Brian W Kunkle,Shubhabrata Mukherjee,Pradeep Natarajan,Adam Naj,Amanda Kuzma,Yi Zhao,Paul K Crane,Alzheimer’s Disease Genetics Consortium,,Hui Lu,Hongyu Zhao
DOI: https://doi.org/10.1038/s41588-019-0345-7
IF: 30.8
2019-01-01
Nature Genetics
Abstract:Transcriptome-wide association analysis is a powerful approach to studying the genetic architecture of complex traits. A key component of this approach is to build a model to impute gene expression levels from genotypes by using samples with matched genotypes and gene expression data in a given tissue. However, it is challenging to develop robust and accurate imputation models with a limited sample size for any single tissue. Here, we first introduce a multi-task learning method to jointly impute gene expression in 44 human tissues. Compared with single-tissue methods, our approach achieved an average of 39% improvement in imputation accuracy and generated effective imputation models for an average of 120% more genes. We describe a summary-statistic-based testing framework that combines multiple single-tissue associations into a powerful metric to quantify the overall gene-trait association. We applied our method, called UTMOST (unified test for molecular signatures), to multiple genome-wide-association results and demonstrate its advantages over single-tissue strategies.