Location tests for biomarker studies: a comparison using simulations for the two-sample case

Autor:	Andreas Ziegler, Markus O. Scheinhardt
Rok vydání:	2012
Předmět:	Location test Health Informatics 01 natural sciences Statistics Nonparametric 010104 statistics & probability 0504 sociology Health Information Management Statistics Econometrics Humans Sample variance Computer Simulation 0101 mathematics Mathematical Computing Mathematics Parametric statistics Advanced and Specialized Nursing Analysis of Variance Models Statistical Homogeneity (statistics) 05 social sciences 050401 social sciences methods Skewness Outlier Linear Models Computerized adaptive testing Monte Carlo Method Biomarkers Type I and type II errors Statistical Distributions
Zdroj:	Methods of information in medicine. 52(4)
ISSN:	2511-705X
Popis:	Summary Background: Gene, protein, or metabolite expression levels are often non-normally distributed, heavy tailed and contain outliers. Standard statistical approaches may fail as location tests in this situation. Objectives: In three Monte-Carlo simulation studies, we aimed at comparing the type I error levels and empirical power of standard location tests and three adaptive tests [O’Gorman, Can J Stat 1997; 25: 269 –279; Keselman et al., Brit J Math Stat Psychol 2007; 60: 267– 293; Szymczak et al., Stat Med 2013; 32: 524 – 537] for a wide range of distributions. Methods: We simulated two-sample scena -rios using the g-and-k-distribution family to systematically vary tail length and skewness with identical and varying variability between groups. Results: All tests kept the type I error level when groups did not vary in their variability. The standard non-parametric U-test per -formed well in all simulated scenarios. It was outperformed by the two non-parametric adaptive methods in case of heavy tails or large skewness. Most tests did not keep the type I error level for skewed data in the case of heterogeneous variances. Conclusions: The standard U-test was a powerful and robust location test for most of the simulated scenarios except for very heavy tailed or heavy skewed data, and it is thus to be recommended except for these cases. The non-parametric adaptive tests were powerful for both normal and non-normal distributions under sample variance homogeneity. But when sample variances differed, they did not keep the type I error level. The parametric adaptive test lacks power for skewed and heavy tailed distributions.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_dedup___::614bc857d4193a73c7d2abd520267e4e https://pubmed.ncbi.nlm.nih.gov/23877617 Zobrazit plný text záznamu