Performing highly parallelized and reproducible GWAS analysis on biobank-scale data-Reference-Cited by-同舟云学术

Performing highly parallelized and reproducible GWAS analysis on biobank-scale data

Published:2023-08-08 Issue: Volume: Page:
ISSN:
Container-title:
language:
Short-container-title:

Author:

Schönherr Sebastian,Schachtl-Riess Johanna,Maio Silvia Di,Filosi Michele,Mark Marvin,Lamina Claudia,Fuchsberger Christian,Kronenberg Florian,Forer Lukas

Abstract

AbstractMotivationGenome-wide association studies (GWAS) in large biobanks are transforming genetic research and enable the detection of novel genotype-phenotype relationships. In the last two decades, over 60,000 genetic associations across thousands of human diseases and traits have been discovered using a GWAS approach. Due to denser genotyping and increasing sample sizes, researchers are increasingly faced with computational challenges when executing GWAS analysis. A reproducible, modular and extensible pipeline with a focus on parallelization is essential to simplify data analysis and to allow researchers to devote their time to other essential tasks such as result interpretation and downstream analysis.ResultsHere we present nf-gwas, a Nextflow pipeline to run biobank-scale GWAS analysis. The pipeline automatically performs numerous pre- and post-processing steps, integrates regression modeling from the REGENIE package and currently supports single-variant, gene-based and interaction testing. nf-gwas also includes an extensive reporting functionality that allows to inspect thousands of phenotypes and navigate interactive Manhattan plots directly in the web browser. The pipeline is extensively tested using the unit-style testing framework nf-test to ensure code maintainability, a crucial requirement in clinical and pharmaceutical settings. Furthermore, we validated the pipeline against published GWAS datasets and benchmarked the pipeline on high-performance computing and cloud infrastructures to provide cost estimations to end users.Availabilitynf-gwas is free available athttps://github.com/genepi/nf-gwas.Contactlukas.forer@i-med.ac.at

Publisher

Cold Spring Harbor Laboratory

Reference15 articles.

1. A brief history of human disease genetics

2. Computationally efficient whole-genome regression for quantitative and binary traits

3. Next-generation genotype imputation service and methods

4. Kassens, JC , Wienbrandt, L and Ellinghaus, D , BIGwas: Single-command quality control and association testing for multi-cohort and biobank-scale GWAS/PheWAS data, Gigascience, 2021;10.

5. H3AGWAS: a portable workflow for genome wide association studies;BMC Bioinformatics,2022

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Nuclear and mitochondrial genetic variants associated with mitochondrial DNA copy number;Scientific Reports;2024-01-24