跳到主要导航 跳到搜索 跳到主要内容

CoMM-S2: A collaborative mixed model using summary statistics in transcriptome-wide association studies

  • Yi Yang
  • , Xingjie Shi
  • , Yuling Jiao
  • , Jian Huang
  • , Min Chen
  • , Xiang Zhou
  • , Lei Sun
  • , Xinyi Lin
  • , Can Yang
  • , Jin Liu*
  • *此作品的通讯作者
  • Shanghai University of Finance and Economics
  • Duke-NUS Medical School
  • Nanjing University of Finance & Economics
  • Zhongnan University of Economics and Law
  • University of Iowa
  • CAS - Academy of Mathematics and System Sciences
  • University of Michigan, Ann Arbor
  • Singapore Clinical Research Institute
  • Agency for Science, Technology and Research, Singapore
  • Hong Kong University of Science and Technology

科研成果: 期刊稿件文章同行评审

摘要

Motivation: Although genome-wide association studies (GWAS) have deepened our understanding of the genetic architecture of complex traits, the mechanistic links that underlie how genetic variants cause complex traits remains elusive. To advance our understanding of the underlying mechanistic links, various consortia have collected a vast volume of genomic data that enable us to investigate the role that genetic variants play in gene expression regulation. Recently, a collaborative mixed model (CoMM) was proposed to jointly interrogate genome on complex traits by integrating both the GWAS dataset and the expression quantitative trait loci (eQTL) dataset. Although CoMM is a powerful approach that leverages regulatory information while accounting for the uncertainty in using an eQTL dataset, it requires individual-level GWAS data and cannot fully make use of widely available GWAS summary statistics. Therefore, statistically efficient methods that leverages transcriptome information using only summary statistics information from GWAS data are required. Results: In this study, we propose a novel probabilistic model, CoMM-S2, to examine the mechanistic role that genetic variants play, by using only GWAS summary statistics instead of individual-level GWAS data. Similar to CoMM which uses individual-level GWAS data, CoMM-S2 combines two models: the first model examines the relationship between gene expression and genotype, while the second model examines the relationship between the phenotype and the predicted gene expression from the first model. Distinct from CoMM, CoMM-S2 requires only GWAS summary statistics. Using both simulation studies and real data analysis, we demonstrate that even though CoMM-S2 utilizes GWAS summary statistics, it has comparable performance as CoMM, which uses individual-level GWAS data.

源语言英语
页(从-至)2009-2016
页数8
期刊Bioinformatics
36
7
DOI
出版状态已出版 - 1 4月 2020
已对外发布

学术指纹

探究 'CoMM-S2: A collaborative mixed model using summary statistics in transcriptome-wide association studies' 的科研主题。它们共同构成独一无二的学术指纹。

引用此