Skip to main navigation Skip to search Skip to main content

An empirical Bayesian ranking method, with applications to high throughput biology

  • Yale University

Research output: Contribution to a Journal (Peer & Non Peer)Articlepeer-review

4 Citations (Scopus)

Abstract

Motivation: In bioinformatics, genome-wide experiments look for important biological differences between two groups at a large number of locations in the genome. Often, the final analysis focuses on a P-value-based ranking of locations which might then be investigated further in follow-up experiments. However, this strategy may result in small effect sizes, with low P-values, being ranked more favorably than larger more scientifically important effects. Bayesian ranking techniques may offer a solution to this problem provided a good prior distribution for the collective distribution of effect sizes is available. Results: We develop an Empirical Bayes ranking algorithm, using the marginal distribution of the data over all locations to estimate an appropriate prior. In simulations and analysis using real datasets, we demonstrate favorable performance compared to ordering P-values and a number of other competing ranking methods. The algorithm is computationally efficient and can be used to rank the entirety of genomic locations or to rank a subset of locations, pre-selected via traditional FWER/FDR methods in a 2-stage analysis.

Original languageEnglish
Pages (from-to)177-185
Number of pages9
JournalBioinformatics
Volume36
Issue number1
DOIs
Publication statusPublished - 1 Jan 2020

Fingerprint

Dive into the research topics of 'An empirical Bayesian ranking method, with applications to high throughput biology'. Together they form a unique fingerprint.

Cite this