Sample size determination for logistic regression revisited

被引:289
|
作者
Demidenko, Eugene [1 ]
机构
[1] Dartmouth Coll Sch Med, Hanover, NH 03755 USA
关键词
case-control study; clinical trial; Fisher information; optimal design; power function; Wald test; Z-score;
D O I
10.1002/sim.2771
中图分类号
Q [生物科学];
学科分类号
07 ; 0710 ; 09 ;
摘要
There is no consensus on the approach to compute the power and sample size with logistic regression. Some authors use the likelihood ratio test; some use the test on proportions; some suggest various approximations to handle the multivariate case. We advocate the use of the Wald test since the Z-score is routinely used for statistical significance testing of regression coefficients. The null-variance formula became popular from early studies, which contradicts modern software, which utilizes the method of maximum likelihood estimation (MLE), when the variance of the MLE is estimated at the MLE, not at the null. We derive general Wald-based power and sample size formulas for logistic regression and then apply them to binary exposure and confounder to obtain a closed-form expression. These formulas are applied to minimize the total sample size in a case-control study to achieve a given power by optimizing the ratio of controls to cases. Approximately, the optimal number of controls to cases is equal to the square root of the alternative odds ratio. Our sample size and power calculations can be carried out online at www.dartmouth.edu/ similar to eugened. Copyright (c) 2006 John Wiley & Sons, Ltd.
引用
收藏
页码:3385 / 3397
页数:13
相关论文
共 50 条