Systematic comparison of ranking aggregation methods for gene lists in experimental results-Reference-Cited by-同舟云学术

Systematic comparison of ranking aggregation methods for gene lists in experimental results

Published:2022-01-10 Issue: Volume: Page:
ISSN:
Container-title:
language:
Short-container-title:

Author:

Wang Bo^ORCID,Law Andy,Regan Tim,Parkinson Nicholas,Cole Joby,Russell Clark D.,Dockrell David H.,Gutmann Michael U.,Baillie J. Kenneth^ORCID

Abstract

AbstractA common experimental output in biomedical science is a list of genes implicated in a given biological process or disease. The results of a group of studies answering the same, or similar, questions can be combined by meta-analysis to find a consensus or a more reliable answer. Ranking aggregation methods can be used to combine gene lists from various sources in meta-analyses. Evaluating a ranking aggregation method on a specific type of dataset before using it is required to support the reliability of the result since the property of a dataset can influence the performance of an algorithm. Evaluation of aggregation methods is usually based on a simulated database especially for the algorithms designed for gene lists because of the lack of a known truth for real data. However, simulated datasets tend to be too small compared to experimental data and neglect key features, including heterogeneity of quality, relevance and the inclusion of unranked lists. In this study, a group of existing methods and their variations which are suitable for meta-analysis of gene lists are compared using simulated and real data. Simulated data was used to explore the performance of the aggregation methods as a function of emulating the common scenarios of real genomics data, with various heterogeneity of quality, noise level, and a mix of unranked and ranked data using 20000 possible entities. In addition to the evaluation with simulated data, a comparison using real genomic data on the SARS-CoV-2 virus, cancer (NSCLC), and bacteria (macrophage apoptosis) was performed. We summarise our evaluation results in terms of a simple flowchart to select a ranking aggregation method for genomics data.

Publisher

Cold Spring Harbor Laboratory

Reference39 articles.

1. A comparative study of rank aggregation methods for partial and top ranked lists in genomic applications;Briefings in bioinformatics,2019

2. Liu, Y.-T. , Liu, T.-Y. , Qin, T. , Ma, Z.-M. , Li, H. : Supervised rank aggregation. In: Proceedings of the 16th International Conference on World Wide Web, pp.481–490 (2007)

3. New approach for understanding genome variations in KEGG

4. The Reactome Pathway Knowledgebase

5. WikiPathways: a multifaceted pathway database bridging metabolomics to other omics research

Cited by 1 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. The genomic landscape of Acute Respiratory Distress Syndrome: a meta-analysis by information content of genome-wide studies of the host response;2024-02-14