Arthur D., Vassilvitskii S. (2007). k-means++: the advantages of careful seeding. Proceedings of the eighteenth annual ACM-SIAM symposium on Discrete algorithms. Society for Industrial and Applied Mathematics Philadelphia, PA, USA. 1027–1035
Abdi H. (2007), Bonferroni and Sidak corrections for multiple comparisons", in N.J. Salkind (ed.): Encyclopedia of Measurement and Statistics. Thousand Oaks, CA: Sage
D'Agostino R.B. and Pearson E.S. (1973), Tests of departure from normality. Empirical results for the distribution of b2 and sqrt(b1). Biometrika, 60, 613-622
D'Agostino R.B., Belanger A., D'Agostino Jr.R B. (1990), A suggestion for using powerful and informative tests of normality. American Statistician, 44, 3 16-321
Agresti A., Coull B.A. (1998), Approximate is better than "exact" for interval estimation of binomial proportions. American Statistics 52: 119-126
Altman D.G., (1998), Confidence intervals for the number needed to treat. BMJ. 317(7168): 1309–1312
Aroian, L. A. (1947), The probability function of the product of two normally distributed variables. Annals of Mathematical Statistics, 18, 265-271.
Bland J.M., Altman D.G. (1999), Measuring agreement in method comparison studies. Statistical Methods in Medical Research 8:135-160.
Anscombe F.J. (1981), Computing in Statistical Science through APL. Springer-Verlag, New York
Armitage P. (1955), Tests for Linear Trends in Proportions and Frequencies. Biometrics. 11 (3): 375–386
Armitage P., Berry G., (1994), Statistical Methods in Medical Research (3rd edition); Blackwell
Armitage P., Colton T., (2009), Encyclopedia of Biostatistics. John Wiley and Sons.
Austin P.C., (2009), The relative ability of different propensity score methods to balance measured covariates between treated and untreated subjects in observational studies. Med Decis Making; 29(6):661-77
Austin P.C., (2011), An Introduction to Propensity Score Methods for Reducing the Effects of Confounding in Observational Studies. Multivariate Behavioral Research 46, 3: 399–424
Barnard G.A. (1989), On alleged gains in power from lower p-values. Statistics in Medicine 8:1469-1477
Baron R. M., Kenny D. A. (1986), The moderator-mediator variable distinction in social psychological research: Conceptual, strategic and statistical considerations. Journal of Personality and Social Psychology, 51, 1173-1182.
Beal S.L. (1987), Asymptotic confidence intervals for the difference between two binomial parameters for use with small samples. Biometrics 43: 941-950
Bender R. (2001), Calculating confidence intervals for the number needed to treat. Controlled Clinical Trials 22:102–110
Benjamini Y. and Hochberg Y. (1995), Controlling the false discovery rate: a practical and powerful approach to multiple testing. Journal of the Royal Statistical Society Series B 57, 289–300
Betty R. Kirkwood and Jonathan A. C. Sterne (2003), Medical Statistics (2nd ed.). Meassachusetts: Blackwell Science, 177\(-\)188, 240\(-\)248
Blanche P., Dartigues J.F., Jacqmin-Gadda H. (2013), Estimating and comparing time-dependent areas under receiver operating characteristic curves for censored event times with competing risks. Statistics in Medicine, 32(30):5381–5397
Bland J.M., Altman D.G. (1986), Statistical methods for assessing agreement between two methods of clinical measurement. Lancet 327 (8476): 307–10
Bland J.M., Altman D.G. (1999), Measuring agreement in method comparison studies. Statistical Methods in Medical Research 8 (2): 135–60
Bowker A.H. (1948), Test for symmetry in contingency tables. Journal of the American Statistical Association, 43, 572-574
Box G. E. , Cox D. R. (1964), An analysis of transformations. Journal of the Royal Statistical Society, Series B 26: 211–252
Breslow N.E., Day N.E. (1980), Statistical Methods in Cancer Research: Vol. I - The Analysis of Case-Control Studies. Lyon: International Agency for Research on Cancer
Breslow N.E. (1996), Statistics in epidemiology: the case-control study', Journal of the American Statistical Association, 91, 14\(-\)28
Brookmeyer R. and Crowley J. (1982a), A confidence interval for the median survival time. Biometrics 38, 29-41
Brown L.D., Cai T.T., DasGupta A. (2001), Interval Estimation for a Binomial Proportion. Statistical Science, Vol. 16, no. 2, 101-133
Brown M.B., Forsythe A. B. (1974a), Robust tests for equality of variances. Journal of the American Statistical Association, 69,364-367
Brown M. B., Forsythe A. B. (1974), The ANOVA and multiple comparisons for data with heterogeneous variances. Biometrics, 30, 719-724
Brown M. B., Forsythe A. B. (1974), The small sample behavior of some statistics which test the equality of several means. Technometrics, 16, 385-389
Brown W. (1910), Some experimental results in the correlation of mental abilities. British Journal of Psychology, 3, 296-322
Hochberg Y. (1988), A sharper Bonferroni procedure for multiple tests of significance. Biometrika 75, 800–803
Chow S.C., Shao J., and Wang H. (2008). Sample Size Calculations in Clinical Research, Second Edition. Chapman and Hall/CRC. Boca Raton, Florida.
Cicchetti D. and Allison T. (1971),A new procedure for assessing reliability of scoring eeg sleep recordings. American Journal EEG Technology, 11, 101-109
Cleveland, W. S. (1979),Robust Locally Weighted Regression and Smoothing Scatterplots. Journal of the American Statistical Association. 74
Clopper C. and Pearson S. (1934), The use of confidence or fiducial limits illustrated in the case of the binomial. Biometrika 26: 404-413
Cochran W.G. (1950), The comparison ofpercentages in matched samples. Biometrika, 37, 256-266
Cochran W.G. (1952), The chi-square goodness-of-fit test. Annals of Mathematical Statistics, 23, 315-345,
Cochran W. G. (1954), Some methods for strengthening the common chi-square tests. Biometrics, 10(4) 17-45 1
Cochran W.G. and Cox G.M. (1957), Experimental designs (2nd 4.). New York: John Wiley and Sons.
Cohen J. (1960), A coefficient of agreement for nominal scales. Educational and Psychological Measurement, 10,3746
Cohen J. (1968), Weighted kappa: nominal scale agreement with provision for scaled disagreement or partial credit. Psychological Bulletin, 70, 213-220
Cohen J. (1988), Statistical Power Analysis for the Behavioral Sciences, Lawrence Erlbaum Associates, Hillsdale, New Jersey
Conover W. J. (1999), Practical nonparametric statistics (3rd ed). John Wiley and Sons, New York
Cox D.R. (1972), Regression models and life tables. Journal of the Royal Statistical Society, B34:187-220
Cramkr H. (1946), Mathematical models of statistics. Princeton, NJ: Princeton University Press.
Cronbach L.J. (1951), Coefficient alpha and the internal structure of tests. Psychometrika, 16(3), 297-334
DeLong E.R., DeLong D.M., Clarke-Pearson D.L., (1988), Comparing the areas under two or more correlated receiver operating curves: A nonparametric approach. Biometrics 44:837-845
Dunn O. J. (1964), Multiple comparisons using rank sums. Technometrics, 6: 241–252
Durbin J. (1951), Incomplete blocks in ranking experiments. British Journal of Statistical Psychology, 4: 85–90
Egger M., Smith G. D., Schneider M., Minder C. (1997), Bias in meta-analysis detected by a simple, graphical test. BMJ, 315(7109):629-634
Epps T.W., Pulley L.B. (1983), A test for normality based on the empirical characteristic function. Biometrika. 1983;70:723–726
Fagerland M. W., Lydersen S., and Laake P. (2013), The McNemar test for binary matched-pairs data: mid-p and asymptotic are better than exact conditional, BMC Med Res Methodol; 13: 91.
Fisher R.A. (1934), Statistical methods for research workers (5th ed.). Edinburgh: Oliver and Boyd
Fisher R.A. (1935), The logic of inductive inference. Journal of the Royal Statistical Society, Series A, 98,39-54
Fisher R.A. (1936), The use of multiple measurements in taxonomic problems. Annals of Eugenics 7 (2): 179–188
Fleiss J.L. (1971), Measuring nominal scale agreement among many raters. Psychological Bulletin, 76 (5): 378–382
Fleiss J.L., Cohen J. (1973), The equivalence of weighted kappa and the intraclass correlation coeffcient as measure of reliability. Educational and Psychological Measurement, 33, 613-619
Fleiss J.L., Levin B., Paik M.C. (2003), Statistical methods for rates and proportions. 3rd ed. (New York: John Wiley) 598-626
Fine J.P., Gray R.J. (1999), A Proportional Hazards Model for the Subdistribution of a Competing Risk. Journal of the American Statistical Association, 94(446):496–509
Freeman G.H. and Halton J.H. (1951), Note on an exact treatment of contingency, goodness of fit and other problems of significance. Biometrika 38:141-149
Freireich E.O., Gehan E., Frei E., Schroeder L.R., Wolman I.J., et al. (1963), The effect of 6-mercaptopmine on the duration of steroid induced remission in acute leukemia. Blood, 21: 699–716
Friedman M. (1937), The use of ranks to avoid the assumption of normality implicit in the analysis of variance. Journal of the American Statistical Association, 32,675-701
Fritz C.O., Morris P.E., Richler J.J.(2012),Effect size estimates: Current use, calculations, and interpretation. Journal of Experimental Psychology: General., 141(1):2–18.
Games P. A., Howell J. F. (1976), Pairwise multiple comparison procedures with unequal n's and/or variances: A Monte Carlo study. Journal of Educational Statistics, 1, 113-125
Gehan E. A. (1965a), A Generalized Wilcoxon Test for Comparing Arbitrarily Singly-Censored Samples. Biometrika, 52:203—223
Gehan E. A. (1965b), A Generalized Two-Sample Wilcoxon Test for Doubly-Censored Data. Biometrika, 52:650—653
Goodman L. A. (1960), On the exact variance of products. Journal of the American Statistical Association, 55, 708-713
Greenhouse S. W., Geisser S. (1959), On methods in the analysis of profile data. Psychometrika, 24, 95–112
Green S.B. (1991), How many subjects does it take to do a regression analysis? Multivariate Behavioral Research, 26, 499-510
Guttman L. (1945), A basic for analyzing test-retest reliabilit. Psychometrika, 10, 255-282
Hanley J.A. i Hajian-Tilaki K.O. (1997), Sampling variability of nonparametric estimates of the areas under receiver operating characteristic curves: an update. Academic radiology 4(1):49-58
Hanley J.A. i McNeil M.D. (1982), The meaning and use of the area under a receiver operating characteristic (ROC) curve. Radiology 143(1):29-36
Hanley J.A. i McNeil M.D. (1983), A method of comparing the areas under receiver operating characteristic curves derived from the same cases. Radiology 148: 839-843
Hanusz Z., Tarasińska J. (2014), On multivariate normality tests using skewness and kurtosis, Colloquium Biometricum 44, 139-148
Henderson, R. (1916), Note on graduation by adjusted average.Transactionsof the Actuarial Society of America, 17:43–48
Henze N., Zirkler B. (1990), A class of invariant consistent tests for multivariate normality. Comm. Statist. Theory Methods. 1990;19:3595–3617
Heagerty P.J., Lumley T., Pepe M.S. (2000), Time-dependent ROC curves for censored survival data and a diagnostic marker. Biometrics, 56(2):337–344
Hochberg Y. (1988), A Sharper Bonferroni Procedure for Multiple Tests of Significance. Biometrika 75 (4): 800–802
Holm S. (1979), A simple sequentially rejective multiple test procedure. Scandinavian Journal of Statistics 6, 65–70
Hotelling H. (1931), The generalization of Student's ratio. Annals of Mathematical Statistics 2 (3): 360–378
Hotelling, H. (1947), Multivariate Quality Control. In C. Eisenhart, M. W. Hastay, and W. A. Wallis, eds. Techniques of Statistical Analysis. New York: McGraw-Hill
Hotelling H. (1951), A generalized t 2 test and measurement of multivariate dispersion. Proceedings of the Second Berkeley Symposium on Mathematical Statistics and Probability 1: 23–41
Huynh H., Feldt L. S. (1976), Estimation of the Box correction for degrees of freedom from sample data in randomized block and split=plot designs. Journal of Educational Statistics, 1, 69–82
Iman R. L., Davenport J. M. (1980), Approximations of the critical region of the friedman statistic, Communications in Statistics 9, 571–595
Jarque C. M., Bera A. K., (1987)., A test for normality of Observations and Regression Residuals, International Statistical Review 55, 163-172
Jones M. C., Marron J. S., Sheather S. J., (1996)., A brief survey of bandwidth selection for density estimation. J. Amer. Statist. Assoc. 91 401–407
Jonckheere A. R. (1954), A distribution-free k-sample test against ordered alternatives. Biometrika, 41: 133–145
Kaplan E.L., Meier P. (1958), Nonparametric estimation from incomplete observations. Journal of the American Statistical Association, 53:457-481
Kendall M.G. (1938), A new measure of rank correlation. Biometrika, 30, 81-93.
Kendall M.G., Babington-Smith B. (1939), The problem of m rankings. Annals of Mathematical Statistics, 10, 275-287
Kleinbaum D. G., Klein M. (2005), Survival Analysis: A Self-Learning Text, Second Edition (Statistics for Biology and Health)
Kolmogorov A.N. (1933), Sulla deterrninazione empirica di una legge di distribuzione. Giornde1l'Inst. Ital. degli. Art., 4, 89-91
Kruskal W.H. (1952), A nonparametric test for the several sample problem. Annals of Mathematical Statistics, 23, 525-540
Kruskal W.H., Wallis W.A. (1952), Use of ranks in one-criterion variance analysis. Journal of the American Statistical Association, 47, 583-621
Lancaster H.O. (1961), Significance tests in discrete distributions. Journal of the American Statistical Association 56:223-234
Lawley D. N. (1938), A generalization of Fisher’s z-test. Biometrika 30: 180–187
Lee E. T., Wang J. W. (2003), Statistical Methods for Survival Data Analysis, ed. third, Wiley
Lenth, R. V., (2001), Some Practical Guidelines for Effective Sample Size Determination. The American Statistician, 55(3), 187-193
Levene H. (1960), Robust tests for the equality of variance. In I. Olkin (Ed.) Contributions to probability and statistics (278-292). Palo Alto, CA: Stanford University Press
Liddell F.D.K. (1983) Simplified exact analysis of case-referent studies; matched pairs; dichotomous exposure. Journal of Epidemiology and Community Health; 37:82-84.
Lilliefors H.W. (1967), On the Kolmogorov-Smimov test for normality with mean and variance unknown. Journal of the American Statistical Association, 62,399-402
Lilliefors H.W. (1969), On the Kolmogorov-Smimov test for the exponential distribution with mean unknown. Journal of the American Statistical Association, 64,387-389
Lilliefors H.W. (1973), The Kolmogorov-Smimov and other distance tests for the gamma distribution and for the extreme-value distribution when parameters must be estimated. Department of Statistics, George Washington University, unpublished manuscript
Lloyd S. P. (1982), Least squares quantization in PCM. IEEE Transactions on Information Theory 28 (2): 129–137
Lund R.E., Lund J.R. (1983), Algorithm AS 190, Probabilities and Upper Quantiles for the Studentized Range. Applied Statistics; 34
Mahalanobis P. C. (1930), On tests and measures of group divergence. Journal of the Asiatic Society of Bengal 26: 541–588
Mahalanobis P. C. (1936), On the generalized distance in statistics. National Institute of Science of India 12: 49–55
Mann H. and Whitney D. (1947), On a test of whether one of two random variables is stochastically larger than the other. Annals of Mathematical Statistics, 1 8 , 5 0 4
Mantel N. and Haenszel W. (1959), Statistical aspects of the analysis of data from retrospective studies of disease. Journal of the National Cancer Institute, 22,719-748
Mantel N. (1963), Chi-square tests with one degree of freedom: Extensions of the Mantel-Haenszel procedure. J. Am. Statist. Assoc., 58, 690-700
Mantel N. (1966), Evaluation of Survival Data and Two New Rank Order Statistics Arising in Its Consideration. Cancer Chemotherapy Reports, 50:163—170
Marascuilo L.A. and McSweeney M. (1977), Nonparametric and distribution-free method for the social sciences. Monterey, CA: Brooks/Cole Publishing Company
Mardia K. V. (1970), Measures of multivariate skewness and kurtosis with applications, Biometrica 57, 519-530
Mardia K. V. (1974), Applications of some measuresof multivariate skewness and kurtosis for testing normality and robustness studies, Sankhay B 36, 115-128
Mauchly J. W. (1940), Significance test for sphericity of n-variate normal population. Annals of Mathematical Statistics, 11, 204-209.
McNemar Q. (1947), Note on the sampling error of the difference between correlated proportions or percentages. Psychometrika, 12, 153-157
Mehta C.R. and Patel N.R. (1986), Algorithm 643. FEXACT: A Fortran subroutine for Fisher's exact test on unordered r*c contingency tables. ACM Transactions on Mathematical Software, 12, 154–161
Miettinen O.S. (1985), Theoretical Epidemiology: Principles of Occurrence Research in Medicine. John Wiley and Sons, New York
Miettinen O.S. and Nurminen M. (1985), Comparative analysis of two rates. Statistics in Medicine 4: 213-226
Mimar S.F. (2017), The Mediation Analysis With the Sobel Test and the Percentile Bootstrap, International Journal of Management and Applied Science, Volume-3, Issue-2
Nadaraya, E. A. (1964), On Estimating Regression. Theory of Probability and Its Applications. 9 (1): 141–2
Newcombe R.G. (1998), Interval Estimation for the Difference Between Independent Proportions: Comparison of Eleven Methods. Statistics in Medicine 17: 873-890
Newman S.C.(2001), Biostatistical Methods in Epidemiology. 2nd ed. New York: John Wiley
Normand S.L. T., Landrum M.B., Guadagnoli E., Ayanian J.Z., Ryan T.J., Cleary P.D., McNeil B.J. (2001), Validating recommendations for coronary angiography following an acute myocardial infarction in the elderly: A matched analysis using propensity scores. Journal of Clinical Epidemiology; 54:387–398.
Ogilvie J. C. (1965), Paired comparison models with tests for interaction. Biometrics 21(3): 651-64
Orwin R. G. (1983), A Fail-SafeN for Effect Size in Meta-Analysis. J Educ Behav Stat, 8(2):157-159
Oyeyemi G.M. , Adewara A.A., Adebola F.B. and Salau S.I. (2010), On the Estimation of Power and Sample Size in Test of Independence, Asian Journal of Mathematics and Statistics, 3(3): 139-146
Page E. B. (1963), Ordered hypotheses for multiple treatments: A significance test for linear ranks. Journal of the American Statistical Association 58 (301): 216–30
Peduzzi P., Concato J., Feinstein A.R., Holford T.R. (1995), Importance of events per independent variable in proportional hazards regression analysis. II. Accuracy and precision of regression estimates. Journal of Clinical Epidemiology, 48:1503-1510
Peduzzi P., Concato J., Kemper E., Holford T.R., Feinstein A.R. (1996), A simulation study of the number of events per variable in logistic regression analysis. Journal of Clinical Epidemiology; 49(12):1373-9
Pillai K. C. (1955), Some new test criteria in multivariate analysis. Annals of Mathematical Statistics 26: 117–121
Plackett R.L. (1984), Discussion of Yates' "Tests of significance for 2x2 contingency tables". Journal of Royal Statistical Society Series A 147:426-463
Pratt J.W. and Gibbons J.D. (1981), Concepts of Nonparametric Theory. Springer-Verlag, New York
Robins, J., Breslow, N., and Greenland S. (1986), Estimators of the Mantel–Haenszel variance consistent in both sparse data and large-strata limiting models. Biometrics 42, 311–323
Robins, J., Greenland S. and Breslow, N.E. (1986), A general estimator for the variance of the Mantel–Haenszel odds ratio. American Journal of Epidemiology 124, 719–723
Rosenbaum P.R., Rubin D.B. (1983a), The central role of the propensity score in observational studies for causal effects. Biometrika; 70:41–55
Rosenthal R. (1979), The "file drawer problem" and tolerance for null results. Psychological Bulletin, 5, 638-641
Rothman K.J., Greenland S., Lash T.L. (2008), Modern Epidemiology, 3rd ed. (Lippincott Williams and Wilkins) 221\(-\)225
Roy S. N. (1939), p-statistics or some generalizations in analysis of variance appropriate to multivariate problems. Sankhya 4: 381–396
Royston P. (1992), Approximating the Shapiro–Wilk W-test for non-normality". Statistics and Computing 2 (3): 117–119
Royston P. (1993b), A toolkit for testing for non-normality in complete and censored samples. Statistician 42: 37–43
Rufibach K. (2010), Assessment of paired binary data; Skeletal Radiology volume 40, pages1–4
Satterthwaite F.E. (1946), An approximate distribution of estimates of variance components. Biometrics Bulletin, 2, 1 10-1 14
Savin N.E. and White K.J. (1977), The Durbin-Watson Test for Serial Correlation with Extreme Sample Sizes or Many Regressors. Econometrica 45, 1989-1996
Schiaparelli, G. V.(1866), Sul modo di ricavare la vera espressione delle leggidelta natura dalle curve empiricae.Effemeridi Astronomiche di Milano perl’Arno, 857:3–56
Scott D. W., (1992), Multivariate Density Estimation. Theory, Practice and Visualization. New York: Wiley.
Shapiro S.S. and Wilk M.B. (1965), An analysis of variance test for normality (complete samples). Biometrika 52 (3–4): 591–611
Sheather S.J. (2009), A modern approach to regression with R. New York, NY: Springer
Shrout P.E., and Fleiss J.L (1979), Intraclass correlations: uses in assessing rater reliability. Psychological Bulletin, 86, 420-428
Šidák Z. K. (1967), Rectangular Confidence Regions for the Means of Multivariate Normal Distributions. Journal of the American Statistical Association, 62 (318): 626–633
Silverman B. W., (1986), Density estimation for statistics and data analysis, London: Chapman and Hall
Skillings J.H., Mack G.A. (1981) On the use of a Friedman-type statistic in balanced and unbalanced block designs. Technometrics, 23:171–177
Sobel M. E. (1982). Asymptotic confidence intervals for indirect effects in structural equation models. Sociological Methodology 13: 290–312
Spearman C. (1910), Correlation calculated from faulty data. British Journal of Psychology, 3, 271-295
Tamhane A. C. (1977), Multiple comparisons in model I One-Way ANOVA with unequal variances. Communications in Statistics, A6 (1), 15-32
Tarone R. E., Ware J. (1977), On distribution-free tests for equality of survival distributions. Biometrica, 64(1):156-160
Tarone R.E. (1985), On heterogeneity tests based on efficient scores. Biometrika 72, 91–95
Terpstra T. J. (1952), The asymptotic normality and consistency of Kendall's test against trend, when ties are present in one ranking. Indagationes Mathematicae, 14: 327–333
Terrell G. R. (1990), The maximal smoothing principle in density estimation. Journal of the American Statistical Association 85, 470–477
Terrell G.R., Scott D. W. (1985), Oversmoothed nonparametric density estimates. Journal of the American Statistical Association 80, 209-214
Thode H. C. (2002), Testing For Normality. CRC Press; 2002. 506 s.
Volinsky C.T., Raftery A.E. (2000) , Bayesian information criterion for censored survival models. Biometrics, 56(1):256–262
Wallenstein S. (1997), A non-iterative accurate asymptotic confidence interval for the difference between two Proportions. Statistics in Medicine 16: 1329-1336
Wallis W.A. (1939), The correlation ratio for ranked data. Journal of the American Statistical Association, 34,533-538
Watson, G. S. (1964), Smooth regression analysis. Sankhyā: The Indian Journal of Statistics, Series A. 26 (4): 359–372
Welch B. L. (1951), On the comparison of several mean values: an alternative approach. Biometrika 38: 330–336
Wilcoxon F. (1945), Individual comparisons by ranking methods. Biometries, 1,80-83
Wilcoxon F. (1945), Individual comparisons by ranking methods. Biometries, 1, 80-83
Wilcoxon F. (1949), Some rapid approximate statistical procedures. Stamford, CT: Stamford Research Laboratories, American Cyanamid Corporation
Wilcoxon F. (1949), Some rapid approximate statistical procedures. Stamford, CT: Stamford Research Laboratories, American Cyanamid Corporation
Wilcoxon F. (1949), Some rapid approximate statistical procedures. Stamford, CT: Stamford Research Laboratories, American Cyanamid Corporation
Wilson E.B. (1927), Probable Inference, the Law of Succession, and Statistical Inference. Journal of the American Statistical Association: 22(158):209-212
Wilks S.S. (1932), Certain generalizations in the analysis of variance. Biometrika 24: 471–494
Yates F. (1934), Contingency tables involving small numbers and the chi-square test. Journal of the Royal Statistical Society, 1,2 17-235
Youden W.J. (1950), Index for rating diagnostic tests. Cancer. 3: 32–35
Kubiak K.B., Więckowska B., Jodłowska-Siewert E., Guzik P. (2024), Visualising and quantifying the usefulness of new predictors stratified by outcome class: The U-smile method. PLOS ONE 19(5): e0303276
Więckowska B., Kubiak K.B., Guzik P. (2025), Evaluating the three-level approach of the U-smile method for imbalanced binary classification. PLOS ONE 20(4): e0321661
Więckowska B., Guzik P. (2026), Usmile likelihood evaluation provides robust threshold free assessment of binary classification models for balanced and imbalanced datasets. Scientific Reports 16: 10000
Yule G. (1900), On the association of the attributes in statistics: With illustrations from the material ofthe childhood society, and c. Philosophical Transactions of the Royal Society, Series A, 194,257-3 19
Zar J. H., (2010), Biostatistical Analysis (Fifth Edition). Pearson Educational
Zweig M.H., Campbell G. (1993), Receiver-operating characteristic (ROC) plots: a fundamental evaluation tool in clinical medicine. Clinical Chemistry 39:561-577