SSizer: Determining the Sample Sufficiency for Comparative Biological Study

Fengcheng Li, Ying Zhou, Xiaoyu Zhang, Jing Tang, Qingxia Yang, Yang Zhang, Yongchao Luo, Jie Hu, Weiwei Xue, Yunqing Qiu, Qiaojun He, Bo Yang (+1 others)
2020 Journal of Molecular Biology  
Comparative biological studies typically require plenty of samples to ensure full representation of the given problem. A frequently-encountered question is how many samples are sufficient for a particular study. This question is traditionally assessed using the statistical power, but it alone may not guarantee the full and reproducible discovery of features truly discriminating biological groups. Two new types of statistical criteria have thus been introduced to assess sample sufficiency from
more » ... fferent perspectives by considering diagnostic accuracy and robustness. Due to the complementary nature of these criteria, a comprehensive evaluation based on all criteria is necessary for achieving a more accurate assessment. However, no such tool is available yet. Herein, an online tool SSizer (https://idrblab.org/ssizer/) was developed and validated to enable the assessment of the sample sufficiency for a user-input biological dataset, and three statistical criteria were adopted to achieve comprehensive and collective assessment. A sample simulation based on a user-input dataset was performed to expand the data and then determine the sample size required by the particular study. In sum, SSizer is unique for its ability to comprehensively evaluate whether the sample size is sufficient and determine the required number of samples for the user-input dataset, which, therefore, facilitates the comparative and OMIC-based biological studies.
doi:10.1016/j.jmb.2020.01.027 pmid:32044343 fatcat:dpxz4mwuyjdmze6ercivlq5guq