Aug 2026· PLoS ONE· Vol 21· 0 citations· 50 references
Medicine
TL;DR
A systematic and updated benchmarking framework for DE analysis is outlined, emphasizing a balance between accuracy and consistency, and a “BaGua (eight trigrams)” map summarizing the multi-dimensional performances of methods is provided.
Abstract
Differential expression (DE) analysis is probably the most prevalent task for transcriptomic studies. However, recent technological advances have seen a revival of methodological interest in DE algorithms. In this study, we performed a comprehensive updated comparative study of 12 representative DE methods using 80 simulated and real datasets. We assessed the adaptability of these methods across varying sample sizes and diverse data scenarios. This evaluation compiled a six-dimensional overview of key properties: detection accuracy, sensitivity at a low false discovery rate, false positives, stability, robustness to outliers, and robustness under noisy conditions. Strikingly, no single methods outperformed others across all evaluation criteria and sample sizes, emphasizing data-specific and scenario-specific method choice. At the widely adopted small-sample size of n = 3, ABSSeq generally outperformed other methods. As sample size increased to n = 5, the sensitivity of DESeq2 and two edgeR v4 algorithms (QLF slightly better than LRT) also raise up under a stringent false-positive control. DESeq had even fewer false positives than DESeq2, at the price of reduced sensitivity. In terms of robustness, Wilcoxon and ROTS are robust to noises for small sample sizes. Moreover, Wilcoxon is also robust to outliers, together with several other methods (ABSSeq, voom, and T.test). NBPSeq and most methods had a good stability even at small sample sizes, except three methods (ROTS, DSS, and T.test). For larger sample sizes (n > 30), all methods performed much better. Finally, we provided a “BaGua (eight trigrams)” map summarizing the multi-dimensional performances of methods, as well as a tree diagram guiding practical method selection. Together, this study outlines a systematic and updated benchmarking framework for DE analysis, emphasizing a balance between accuracy and consistency.
Accurate normalization is essential for differential expression analysis of RNA-sequencing data. Popular normalization methods such as the median-of-ratios and trimmed mean of M-values do not leverage information from the experimental design. This may be inefficient in experiments with large-scale systematic expression...
Todd Pocuca, Guillaume Paré, B. M. Bolker· bioRxiv· 0 citations
Differential expression/abundance analyses are commonplace in studies employing high-throughput sequencing (HTS); however different tools often fail to return comparable results when applied to the same dataset. Most tools employ normalisations to attempt to correct for technical variation in the count data. Previously...
S. D. Dos Santos, A. Murariu, Justin D. Silverman et al.· PLoS Computational Biology· 0 citations
Computer methods that simulate gene knockout experiments from single-cell RNA sequencing data are increasingly popular, but researchers lack guidance on which method to choose, so preliminary guidance for method selection is established, including cross-pathway validation, direction-aware benchmarking, and minimum data...
Shixiang Wu, Gang Hu, Zhang-Quan Yang et al.· bioRxiv· 0 citations
Feature selection is critical for resolving cell-type heterogeneity in single-cell RNA sequencing (scRNA-seq). DUBStepR (Determining the Underlying Basis using Stepwise Regression) is a widely used gene selection method for scRNA-seq designed to identify feature genes that maximize cell-type separation. DUBStepR has be...
Machine learning-based annotation methods are increasingly used to assess the pathogenicity of genetic variants, but their performance at prioritizing variants for gene-level association testing remains poorly characterized. Here, to better understand and optimize for this use case, we assess variant annotations from...
Matthew Aguirre, Flaviyan Jerome Irudayanathan, M. Crow et al.· BMC Genomics· 0 citations
To support reproducible best practice, a simple command set is provided for selecting and documenting study-appropriate backgrounds and for assessing sensitivity of GO Biological Process results to the chosen universe.
Brian Timoney, P. Guasoni, Komal Zade et al.· bioRxiv· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.