SSizer: Determining the Sample Sufficiency for Comparative Biological Study

2020 
Abstract Comparative biological studies typically require plenty of samples to ensure full representation of the given problem. A frequently-encountered question is how many samples are sufficient for a particular study. This question is traditionally assessed using the statistical power, but it alone may not guarantee full and reproducible discovery of features truly discriminating biological groups. Two new types of statistical criteria have thus been introduced to assess sample sufficiency from different perspectives by considering diagnostic accuracy and robustness. Due to the complementary nature of these criteria, a comprehensive evaluation based on all criteria is necessary for achieving more accurate assessment. However, no such tool is available yet. Herein, an online tool SSizer ( https://idrblab.org/ssizer/ ) was developed and validated to enable the assessment of the sample sufficiency for a user-input biological dataset, and three statistical criteria were adopted to achieve comprehensive and collective assessment. A sample simulation based on user-input dataset was performed to expand the data and then determine the sample size required by particular study. In sum, SSizer is unique for its ability to comprehensively evaluate whether the sample size is sufficient and determine the required number of samples for user-input dataset, which therefore facilitate the comparative and OMIC-based biological studies.
    • Correction
    • Source
    • Cite
    • Save
    • Machine Reading By IdeaReader
    55
    References
    18
    Citations
    NaN
    KQI
    []