Sample size determination for logistic regression revisited
Top Cited Papers
- 6 December 2006
- journal article
- research article
- Published by Wiley in Statistics in Medicine
- Vol. 26 (18) , 3385-3397
- https://doi.org/10.1002/sim.2771
Abstract
There is no consensus on the approach to compute the power and sample size with logistic regression. Some authors use the likelihood ratio test; some use the test on proportions; some suggest various approximations to handle the multivariate case. We advocate the use of the Wald test since the Z‐score is routinely used for statistical significance testing of regression coefficients. The null‐variance formula became popular from early studies, which contradicts modern software, which utilizes the method of maximum likelihood estimation (MLE), when the variance of the MLE is estimated at the MLE, not at the null. We derive general Wald‐based power and sample size formulas for logistic regression and then apply them to binary exposure and confounder to obtain a closed‐form expression. These formulas are applied to minimize the total sample size in a case–control study to achieve a given power by optimizing the ratio of controls to cases. Approximately, the optimal number of controls to cases is equal to the square root of the alternative odds ratio. Our sample size and power calculations can be carried out online at www.dartmouth.edu/∼eugened. Copyright © 2006 John Wiley & Sons, Ltd.Keywords
This publication has 22 references indexed in Scilit:
- On Power and Sample Size Calculations for Likelihood Ratio Tests in Generalized Linear ModelsBiometrics, 2000
- Sixteen S‐squared over D‐squared: A relation for crude sample size estimatesStatistics in Medicine, 1992
- Some Surprising Results about Covariate Adjustment in Logistic Regression ModelsInternational Statistical Review, 1991
- Power/Sample Size Calculations for Generalized Linear ModelsBiometrics, 1988
- Calculating Sample Sizes in the Presence of Confounding VariablesJournal of the Royal Statistical Society Series C: Applied Statistics, 1986
- On the Use of Wald's Test in Exponential FamiliesInternational Statistical Review, 1985
- The Design of Case-Control Studies: The Influence of Confounding and Interaction EffectsInternational Journal of Epidemiology, 1984
- Sample Size for Logistic Regression with Small Response ProbabilityJournal of the American Statistical Association, 1981
- Logistic disease incidence models and case-control studiesBiometrika, 1979
- Wald's Test as Applied to Hypotheses in Logit AnalysisJournal of the American Statistical Association, 1977