Estimating the Number and Effect Sizes of Non-null Hypotheses
Jennifer Brennan, Ramya Korlakai Vinayak, Kevin Jamieson
Abstract
We study the problem of estimating the distribution of effect sizes (the mean of the test statistic under the alternate hypothesis) in a multiple testing setting. Knowing this distribution allows us to calculate the power (type II error) of any experimental design. We show that it is possible to estimate this distribution using an inexpensive pilot experiment, which takes significantly fewer samples than would be required by an experiment that identified the discoveries. Our estimator can be used to guarantee the number of discoveries that will be made using a given experimental design in a future experiment. We prove that this simple and computationally efficient estimator enjoys a number of favorable theoretical properties, and demonstrate its effectiveness on data from a gene knockout experiment on influenza inhibition in Drosophila. Estimating the Number and Effect Sizes of Non-null Hypotheses Our estimator indicates discoveries exist… Which of 10,000 Drosophila genes inhibit virus growth? Option 1 Full Experiment Find all genes that inhibit virus replication by at least 2x, as measured by a fluorescent reporter Replicates:
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 72036cde-c338-4148-aa28-51a06631f0e2Cited by top-tier papers1
Ask how each one uses itRelated papers
- Data Amplification: Instance-Optimal Property EstimationYi Hao, Alon OrlitskyICML 2020 · 23 citations
- DiscoBAX: Discovery of optimal intervention sets in genomic experiment designClare Lyle, Arash Mehrjou, Pascal Notin, Andrew Jesson et al.ICML 2023 · 16 citations
- Effect Size Estimation for Duration Recommendation in Online Experiments: Leveraging Hierarchical Models and Objective Utility ApproachesYu Liu, Runzhe Wan, James McQueen, Doug Hains et al.AAAI 2024 · 1 citation
- How Much is Unseen Depends Chiefly on Information About the SeenSeongmin Lee, Marcel BöhmeICLR 2025
- Asymptotically Optimal and Computationally Efficient Average Treatment Effect Estimation in A/B testingVikas Deep, Achal Bassamboo, Sandeep K. JunejaICML 2024 · 1 citation
