Methods Inf Med 2013; 52(04): 351-359
DOI: 10.3414/ME12-02-0014
Focus Theme – Original Articles
Schattauer GmbH

Location Tests for Biomarker Studies: A Comparison Using Simulations for the Two-sample Case[*]

M. O. Scheinhardt
1   Institut für Medizinische Biometrie und Statistik, Universität zu Lübeck, Universitätsklinikum Schleswig-Holstein, Campus Lübeck, Lübeck, Germany
,
A. Ziegler
1   Institut für Medizinische Biometrie und Statistik, Universität zu Lübeck, Universitätsklinikum Schleswig-Holstein, Campus Lübeck, Lübeck, Germany
2   Zentrum für Klinische Studien, Universität zu Lübeck, Lübeck, Germany
3   Deutsches Zentrum für Herz-Kreislaufforschung, Standort Hamburg/Kiel/Lübeck, Lübeck, Germany
› Author Affiliations
Further Information

Publication History

received: 03 December 2012

accepted: 09 June 2013

Publication Date:
20 January 2018 (online)

Summary

Background: Gene, protein, or metabolite expression levels are often non-normally distributed, heavy tailed and contain outliers. Standard statistical approaches may fail as location tests in this situation.

Objectives: In three Monte-Carlo simulation studies, we aimed at comparing the type I error levels and empirical power of standard location tests and three adaptive tests [O’Gorman, Can J Stat 1997; 25: 269 –279; Keselman et al., Brit J Math Stat Psychol 2007; 60: 267– 293; Szymczak et al., Stat Med 2013; 32: 524 – 537] for a wide range of distributions.

Methods: We simulated two-sample scena -rios using the g-and-k-distribution family to systematically vary tail length and skewness with identical and varying variability between groups.

Results: All tests kept the type I error level when groups did not vary in their variability. The standard non-parametric U-test per -formed well in all simulated scenarios. It was outperformed by the two non-parametric adaptive methods in case of heavy tails or large skewness. Most tests did not keep the type I error level for skewed data in the case of heterogeneous variances.

Conclusions: The standard U-test was a powerful and robust location test for most of the simulated scenarios except for very heavy tailed or heavy skewed data, and it is thus to be recommended except for these cases. The non-parametric adaptive tests were powerful for both normal and non-normal distributions under sample variance homogeneity. But when sample variances differed, they did not keep the type I error level. The parametric adaptive test lacks power for skewed and heavy tailed distributions.

* Supplementary material published on our website www.methods-online.com