Scalable and Robust Regression Methods for Phenome-Wide Association Analysis on Large-Scale Biobank Data

Wenjian Bi; Seunggeun Lee

ArticleOPEN ACCESS

Scalable and Robust Regression Methods for Phenome-Wide Association Analysis on Large-Scale Biobank Data

Frontiers in Genetics

DOI: 10.3389/fgene.2021.682638

6Citations

18Readers

Abstract

With the advances in genotyping technologies and electronic health records (EHRs), large biobanks have been great resources to identify novel genetic associations and gene-environment interactions on a genome-wide and even a phenome-wide scale. To date, several phenome-wide association studies (PheWAS) have been performed on biobank data, which provides comprehensive insights into many aspects of human genetics and biology. Although inspiring, PheWAS on large-scale biobank data encounters new challenges including computational burden, unbalanced phenotypic distribution, and genetic relationship. In this paper, we first discuss these new challenges and their potential impact on data analysis. Then, we summarize approaches that are scalable and robust in GWAS and PheWAS. This review can serve as a practical guide for geneticists, epidemiologists, and other medical researchers to identify genetic variations associated with health-related phenotypes in large-scale biobank data analysis. Meanwhile, it can also help statisticians to gain a comprehensive and up-to-date understanding of the current technical tool development.

Author supplied keywords

Cite

CITATION STYLE

APA

Bi, W., & Lee, S. (2021, June 15). Scalable and Robust Regression Methods for Phenome-Wide Association Analysis on Large-Scale Biobank Data. Frontiers in Genetics. Frontiers Media SA. https://doi.org/10.3389/fgene.2021.682638

Scalable and Robust Regression Methods for Phenome-Wide Association Analysis on Large-Scale Biobank Data

Abstract

Author supplied keywords

Cite

Register to see more suggestions