Abstract
Reaction time (RT) is one of the most common types of measure used in experimental psychology. Its distribution is not normal (Gaussian) but resembles a convolution of normal and exponential distributions (Ex-Gaussian). One of the major assumptions in parametric tests (such as ANOVAs) is that variables are normally distributed. Hence, it is acknowledged by many that the normality assumption is not met. This paper presents different procedures to normalize data sampled from an Ex-Gaussian distribution in such a way that they are suitable for parametric tests based on the normality assumption. Using simulation studies, various outlier elimination and transformation procedures were tested against the level of normality they provide. The results suggest that the transformation methods are better than elimination methods in normalizing positively skewed data and the more skewed the distribution then the transformation methods are more effective in normalizing such data. Specifically, transformation with parameter lambda -1 leads to the best results.
INTRODUCTION
Reaction times (RTs) have been a privileged measure of behavior in experimental psychology allowing an estimation of the duration of cognitive processes and inference of the likely cognitive process (see ). Hence, their understanding and proper analysis is essential. It is known that reaction time data are positively skewed, and therefore are not normally distributed. As Olivier and Norberg (2010) argue, commonly used statistical tests are not appropriate for the analysis of RT data since RTs are (in most cases) non-normally distributed. Yet, most researchers rely on parametric tests (primarily ANOVA) to analyze reaction times data despite these tests assumptions are not met with RT data. More specifically, they require variables to be normally distributed within conditions and have homogeneous variances between conditions in order to give unbiased results (e.g., ; ). Even small violations of those assumptions can lead to biased results from the tests (see Wilcox, 1998). To meet these conditions, many researchers transform the data and/or search for maverick data points. The aim of this paper is to compare various procedures that assist in normalizing data via power transformations and outlier elimination procedures.
REACTION TIME DISTRIBUTIONS AND METHODS TO DEAL WITH OUTLIERS
Reaction time distributions are characterized by a positive skew. Many explanations have been proposed to explain this near universal finding (with one exception; , who found symmetrical RT distributions). The first explanation from , argued that observable RTs are caused by two processes operating in succession. The first is a central decision mechanism whose distribution is highly skewed (exponential distribution). This mechanism is related to an accumulation of information processes whose activation times have a rate of accumulation τ ms-1. These assumptions are based on neurological studies of single cell firing patterns (see e.g., ) showing that for a small threshold, the resulting distribution is very skewed and well-described by an exponential distribution. The second process is responsible for response selection and motor execution. This second process is presumably affected by many factors and therefore (owing to the central limit theorem) results in a normal distribution. The sum of these two sets of time has a distribution described by a convolution of a Gaussian distribution and an exponential distribution known as an Ex-Gaussian distribution. , Ratcliff and Murdock (1976), , and , among others, fitted this distribution to RTs and found a generally good fit.
In other words, a simple cognitive task starts at the level of the perceptual processes. Light travels to the retina (negligible time), activating the cones and rods on the retinae and transducing the signal through the visual brain areas (V1, V2, etc.). Following perception and up to a semantically meaningful percept, there is the decision process presumably occurring in the frontal lobes. A decision is then followed by activations sent for response selection and down to the motor areas and spinal cord triggering a muscular response, which puts pressure on a response key. Overall, a simple decision involves a chain of signal transformation through a dozen specialized brain areas each adding to the observed latency. The total processing chain can be subdivided into three stages: perception, decision and response selection, and motor response. Based on the assumption that the time taken by each brain area adds up to the total response time observed by the apparatus, and knowing that manipulating the difficulty of the decision without altering the perceptibility of the stimuli and without altering the motor response complexity can affect skew, it can be hypothesized that (1) perception processes add up to a total perception time; (2) response selection and motor response processes add up to a total response time, and (3) the balance leads to a decision time. Although the time taken by the perceptual processes are unknown, owing to the Central limit theorem, if multiple processes with unknown times are added to obtain a total time then the resulting perception time should be normally distributed. This same principle applies to the response selection and motor response stage.
In recent years, though, some authors have questioned the additivity assumptions. Under an alternative view of the chain of processing, the brain areas send activations continuously and related areas react when a critical amount of activations have been received. Hence, each area is not operating in isolation. Thus, violating the independence of operation of each sub-process implicit in the additivity assumption. Theorems analogous to the Central limit theorem suggest that the resulting perception time should be log-normally distributed in this scenario (Ulrich and Miller, 1993; Mouri, 2013). Finally, as the decision processes are based on just a few sub-processes, asymptotic theorems cannot be invoked and this stage should preserve its highly skewed characteristic at the level of latency.
Other explanations have subsequently been proposed to explain the skew in RTs. Ulrich and Miller (1993) suggested that response times may be caused by a cascade of events (following , cascade model). This model predicts a distribution called the Log-Normal (also see West and Shlesinger, 1990) whose shape is indistinguishable from the EGd (Chechile, personal communication). Raab (1962), followed by and Pike (1973), instead proposed a race model where brain signals compete with each other to be the first to trigger a response (recent documentation includes Rouder, 2000; Miller and Ulrich, 2003; ). These models all suggest that the Weibull distribution should be the distribution of RTs (see also Schwarz, 2001 for a variation of this idea). The Weibull and Ex-Gaussian/Log-Normal can be in principle distinguished, but this requires a lot of RTs per conditions (more than 100), uncontaminated by practice effects, fatigue effects, etc. Nevertheless, the true distribution of RT may also (more likely) be none of the above.
In what follows, we assume that the EGd is a convenient way to characterize RTs much like statisticians assume the normal distribution. Furthermore, the literature indicates that the EGd is the distribution most broadly explored. A comprehensive characterization of EGds can be found in . The parameters of an EGd are represented by a mean (μ), a SD (σ), and an exponential factor (τ). The mean and the SD represent times from the normally distributed stage of processing, whereas the exponential factor represents times from the exponentially distributed stage of processing (see ). The mean of an EGd can be inferred from its parameters, as well as its standard variation. The EGd’s SD and exponential factor can be used to estimate its third (skewness) and fourth (kurtosis) central moments (see Figure 1).
FIGURE 1
It is a common practice in psychological research that measures RT data to deal with observations which fall very far from a group’s mean. Such observations are possibly the product of participants’ lack of attention (very long RTs) or overly fast guesswork (very fast RTs). However, it may also be caused by random fluctuations in internal thresholds. These outliers deeply affect the estimation of the data’s central tendency (see Whelan, 2008;
Under the 2 SD procedures, researchers remove observations which fall ± 2 SD from a participant’s mean in a particular condition (see Ratcliff, 1993; for an example of this application see
The SD cut-off is the most commonly used procedure to deal with outliers in RT research. However, advances in statistics suggest the use of more robust methods to deal with outliers. One such approach derives from multivariate outlier detection methods and is called the minimum covariance determinant (MCD) method. This method aims to estimate the best subset of normally distributed points in a data set which are clustered in an ellipsoid with the smallest volume (or minimum covariance matrix). The computations of the MCD rely on Mahalanobis distances and robust estimators of multivariate location (see Rousseeuw and van Driessen, 1999). Although the MCD method is primarily designed to deal with multivariate data, it does not preclude it from being applied to univariate data.
Another approach to deal with non-normality is data transformations. With this procedure all observations are retained but they are re-expressed using a different, non-linear, scale that improves normality of the data (see Osborne, 2002; Olivier and Norberg, 2010). RT data can be re-expressed into logarithms (for an example of this application see
A simulation study in which the normalization power of the Box–Cox transformations and elimination procedures are tested against a particular type of skewed distributions is yet to be done. As to the Box–Cox method, it would be useful to see how other transformation parameters could improve the normality of EGds. Thus, the intermediate parameters -0.5 – are worth testing since it can be seen as a trade-off between an inverse (i.e., -1) and a logarithmic (i.e., 0) transformation. The present simulation study aims to test the power of these outlier elimination and transformation methods to normalize EGds of different parameters and sample sizes. The results will indicate the most effective methods when dealing with positively skewed distributions.
MATERIALS AND METHODS
VALIDATION OF AN ALTERNATIVE SIMULATION METHOD AND A COMPREHENSIVE APPROACH TO THE ASSESSMENT OF NORMALITY IN NON-NORMAL DISTRIBUTIONS
In order to determine how various outlier elimination and transformation methods can improve the normality of data sampled from EGds, it is necessary to first check the normality of the EGds before applying these methods. A typical approach is to estimate the power of normality tests against non-normal distributions. Under this approach it is traditional to (i) Compute the Critical values (CVs) of one or more normality tests against a N(0,1) of different sample sizes, and to (ii) Use those CVs as cut-off points to reject normality in non-normal distributions of the same sample sizes used in the simulations (see
The power of a normality test relies on the number of times the test correctly rejects normality. We note hereafter the proportion of rejection of normality as PoR. A high PoR (e.g., a PoR close to 1) signifies that the distribution being tested is highly non-normal, whereas a low PoR indicates otherwise (e.g., a PoR close to 0). On the other hand, all tests should show PoRs hovering around 0.05 (as α = 0.05) when tested against a normally distributed set of data regardless of the sample size (e.g., Romão et al., 2010). Such a situation is to be expected since normality tests should have a low probability of incorrectly rejecting the hypothesis that a N(0,1) is normal, and that probability should be close to the nominal level used in the study (the α level). In sum, normality tests should have a low PoR, against normal distributions and have a high power, or a high PoR, against non-normal distributions.
In a recent study,
The present study features an alternative approach in which the p-value associated with a normality test is used.
To validate the p-value approach, a simulation study was performed to test the power of the SW against the same EGds described above, when sample size was 10, and with an alpha level of 0.05. Additionally, the present simulations implement the method proposed by
As suggested above, it is essential to determine the status of non-normal EGds before any outlier elimination or transformation method is applied. As is traditional in simulation studies of normality tests (see Romão et al., 2010;
An alternative method in which the combined results of various normality tests are used is proposed herein. There are approximately 40 different types of normality tests (see Razali and Wah, 2010) that can be categorized as regression/correlations, empirical distribution functions, measure of moments, or a combination of these (see Romão et al., 2010;
FIGURE 2

Mean PoR (and ±1 SD) of a combined set of six normality tests for three EGds when n = 10, 15, 20, 30, and 50. The associated mean p-values for each case, and their ±1 SDs (in parenthesis), are shown in italics and between brackets. p-values below 0.05 are bolded.
This section aims to determine the degree of normality achieved by the outlier elimination and transformation methods described above, on various EGds. The six normality tests described above were used to provide a gage of the average level of normality achieved by the outlier methods. The simulation approach described above, which uses iterations of simulations and estimation of an average, was used in the study. In addition, the p-value approach described above was used to determine the PoRs after the outlier methods were applied to the EGds.
Three different sets of EGds were generated. The parameters were those described above: μ = 300, σ = 20, and τ = 300 (EGd1), μ = 400, σ = 20, and τ = 200 (EGd2), and μ = 500, σ = 20, and τ = 100 (EGd3). Each EGd was generated in four sample sizes: 15, 20, 30, and 502. These parameters represent actual RT data and are taken from the 12 different EGds reported originally by
Thus, for each combination of EGd, sample size, and outlier method, i number of PoR and p-value results were available. The mean results of the PoRs and p-values were submitted to a 3 (types of EGd = EGd1, EGd2, and EGd3) × 4 (sample sizes = 15, 20, 30, and 50) × 2 (outlier methods = transformation and elimination) ANOVA-type statistic (ATS; see Noguchi et al., 2012 for details) in order to determine main effects and interactions. The “type of EGds” was entered in the analysis as the between-subjects factor, while the other factors were entered as the within-subjects factors. The items for the outlier transformation method were the four transformation parameters of the Box–Cox transformations described above, i.e., -1, -0.5, 0, and 0.5 and the items for the outlier elimination method were the four methods discussed above, i.e., the MCD, ± 2 SD, ± 2.5 SD, and ± 3 SD methods. Comparisons of two dependent groups were performed via the Yuen test (Ty; see Wilcox, 2012 for details).
In the particular case of the outlier elimination methods, the proportion of data eliminated (PoE) was estimated in the same way as the PoRs. That is, for each distribution to which an outlier elimination method was applied, the proportion of observations removed by a specific method was computed. Then, an average of PoE was estimated for the total number of simulations and each simulation run was iterated i times. Finally, the mean PoEs across iterations were computed.
Also, for both outlier methods, the mean p-value is reported in order to render more obvious the direct relationship between PoR and p-values. That is, the higher the mean p-value, the lower the PoR, and the lower the mean p-value, the higher the PoR. That is, more chances of normality rejection are paired with mean p-values, signaling non-normality. Figure 3 illustrates the key features of the simulation study.
FIGURE 3

Key components of the simulation study. Specific details can be found in the ‘Materials and Methods’ section. IV, independent variables (BS, between-subjects factors; WS, within subjects factors; D, distributions; n, sample sizes, OM, outlier methods), S, simulation (s, number of simulations; i, number of iterations), DV, dependent variables [M (SD) = means and standard deviations; 6NT, mean across six normality tests; ∀OM, for all outlier methods; ≡EM, only for elimination methods], and SA, statistical analysis (ATS, ANOVA-type statistic).
A NOTE ON THE CHARACTERISTICS OF THE PRESENT SIMULATIONS
All throughout this article the simulation method used in the present study has been depicted so it is worthwhile emphasizing the value of the simulation approach used herein. Canonical simulation studies on normality test report tables of a unique number that represents the power of the test under study, i.e., the proportion of times the test rejected normality (here PoR) at the alpha level chosen. For instance, in an ideal simulation study in which 20’000 simulations are run, it is expected that in 1’000 of them the assumption of normality is incorrectly rejected when tested against a N(0,1), which in turn, gives a power of 0.05 (or a PoR of 0.05). However, if such simulation is run a second, third, or an x number of times, it is likely that the number of N(0,1) distributions flagged as non-normal, varies from 1000 every time. Therefore, giving a PoR of approximately 0.05. Such variation in the outcome can be due to several factors such as the type of RNG used, the use of seeding in the simulations, the statistical package used, and/or simply chance.
There is in fact another issue associated with the study of normality tests that can play a part in the simulation process. When a normality test is used against a N(0,1), a distribution of x number of observations, i.e., the number of simulations, containing the results of the test statistic is formed. CVs are then obtained by calculating the key quantiles of the test statistic’s distribution (e.g., the 95% quantile in positively skewed distributions when alpha is 0.05). However, the CVs found are directly dependent on the computation used to estimate the quantiles and there are various computations involved [for instance, the R software implements nine different quantile computations (see
RESULTS
PROPORTION OF REJECTION AND P-VALUES
The mean PoR and mean p-values corresponding to the transformation of outliers via the Box–Cox transformation parameters are shown in Figure 4. Figure 5 shows the mean PoR and mean p-values for the case of elimination of outliers via the MCD and the ±n SD methods.
FIGURE 4

Mean proportion of rejection (four top panels) and p-values (four bottom panels) of the outlier accommodation procedure via data transformation. Lambda represents the parameter used to perform the transformation. Insets show the mean and SD estimates across distributions types and sample sizes.
FIGURE 5

Mean proportion of rejection (four top panels) and p-values (four bottom panels) of the outlier elimination procedures via the MCD and SD methods. Insets show the mean and SD estimates across distributions types and sample sizes.
The ATS analyses suggested significant main effects of distribution type; sample size and outlier method and their two and three way interactions in both the PoR and p-value analyses (see Table 1). The effect sizes of the three way interactions are shown in Figure 6.
Table 1
| DV | Main effect | Interaction |
|---|---|---|
| PoR | FD(1.94,82.53) = 4165.33, p < 0.001 | FD×O(1.96,∞) = 22335.40, p < 0.001 |
| FO(1,∞) = 31857.81, p < 0.001 | FS×O(2.72,∞) = 883.00, p < 0.001 | |
| Fs(2.54,∞) = 78415.56, p < 0.001 | FD×S (4.90,∞) = 525.59, p < 0.001 FD×O×S (5.17,∞) = 267.56, p < 0.001 | |
| pV | FD(1.93,81.51) = 1681.45, p < 0.001 | FD×O(1.96,∞) = 34226.12, p < 0.001 |
| FO(1,∞) = 74198.60, p < 0.001 | FS×O(2.85,∞) = 502.25, p < 0.001 | |
| Fs(2.82,∞) = 82610.79, p < 0.001 | FD×S(5.50,∞) = 498.73, p < 0.001 FD×O×S(5.57,∞) = 255.15, p < 0.001 | |
| PoE | FD(1.96,83.66) = 4382.35, p < 0.001 | FD×E(5.02,∞) = 165.09, p < 0.001 |
| FE(2.66,∞) = 267304.12, p < 0.001 | FS×E(6.15,∞) = 4529.77, p < 0.001 | |
| Fs(2.82,∞) = 186.08, p < 0.001 | FD×S(5.42,∞) = 12.11, p < 0.001 FD×E×S(11.18,∞) = 35.62, p < 0.001 |
Results of the ANOVA-type statistic (ATS).
DV, dependent variable; PoR, proportion of rejection data; pV, p-value data; PoE, proportion of elimination data; D, distribution type (EGd1, EGd2, and EGd3); O, outlier method (elimination and transformation); S, sample size (15, 20, 30, and 50), and E, elimination method (MCD, ±2 SD, ±2.5 SD, and ±3 SD).
FIGURE 6

Relative treatment effect plot of the three way interaction between distribution type, sample size, and outlier method for the proportion of rejection and p-value analyses.
As shown in the first row in Figure 6, the likelihood of the rejection of normality increases as sample size (i.e., from 15 to 50) and distribution skewness increase (i.e., from EGd3 to EGd1). This result is corroborated by the p-values analyses in that as sample size and distribution skewness increase, the p-values decrease (second row in Figure 6). This is a phenomenon recognized in research on the power of normality tests (see
As Figures 4 and 6 indicate, the transformation methods seem to lead to decreased normality rejection as compared with elimination methods. The larger effects of the former methods over the latter are summarized in Figure 6. As shown in Figure 4, the transformation methods seem to be particularly useful when dealing with highly skewed distributions (i.e., EGd1) in that, across sample sizes, low PoRs, and high p-values were obtained for these distributions after transformation. Specifically, the results indicate that transformations with lambda -1 would seem to provide the best results (see insets in Figure 4). These results are in agreement with past research suggesting that the inverse transformation has a strong normalization effect (see Ratcliff, 1993).
PROPORTION OF ELIMINATION
The mean PoE corresponding to the elimination of outliers via the MCD and ±n SD methods is shown in Figure 7.
FIGURE 7

Mean proportion of elimination of the outlier elimination procedures via the MCD and SD methods. Insets show the mean and SD estimates across distributions types and sample sizes.
The ATS analyses suggested significant effects of distribution type; sample size, and outlier elimination method and their two and three way interactions in the PoE analyses (see Table 1). The effect sizes of the three way interactions are shown in Figure 8.
FIGURE 8

Relative treatment effect plot of the three way interaction between distribution type, sample size, and outlier method for the proportion of elimination analyses.
Figure 5 reports the mean proportion of rejection achieved by each method for each distribution in different sample sizes and their associated p-values. The results suggest that the MCD method seems to lead to lower PoR and higher p-values than the SD methods. However, by comparing Figures 5 and 7, a trade-off between the PoRs (and associated p-values) and the POE associated with each of these methods becomes clear. Thus, while the MCD method leads to the lowest PoRs, it does have the highest POE. On the contrary, the ±3 SD method has low PoE but at the cost of leading to a rather high proportion of normality rejection.
As the relative treatment effect plot in Figure 8 indicates, for all methods, the likelihood of eliminating more data increases as the distribution becomes more skewed; i.e., from EGd 3 to EGd 1. However, while the MCD and ±2 SD’s likelihood of rejecting data appears to reduce as sample size increases, for the remaining methods such a likelihood increases as sample size increases.
In summary, the key result is that the transformation methods are more effective than the elimination methods at normalizing positively skewed distributions. That is, the outlier method had a main effect. Thus, indicating that across distributions types and sample sizes the transformation methods led to lower PoR (MPoR = 0.29, SD = 0.17) and higher p-values (Mp-value = 0.24, SD = 0.09) than the elimination methods (MPoR = 0.42, SD = 0.22; Mp-value = 0.16, SD = 0.09) [PoR: Ty (359) = -14.54, p < 0.001; p-value: Ty (359) = 20.26, p < 0.001].
DISCUSSION
The results of this simulation study suggest that the Box–Cox transformation methods outperform the elimination methods in normalizing positively skewed data and the more skewed the distribution, the more effective the transformation methods in normalizing such type of data. Various ideas need to be discussed in relation to this finding.
The difference between transformations and elimination procedures is that transformations seek to stabilize variance and skewness (see
Applying the methods studied herein to data believed to be non-normal, does not automatically guarantee that the data has met parametric assumptions. That is, it is important to corroborate, via graphical and formal tests, that these assumptions have been reached. Although this is a well-known recommendation, it is rare to find published papers reporting normality or homogeneity tests in order to justify the use of parametric analyses. It is therefore important that whatever method is used to filter data, formal normality, and homogeneity tests are reported in order to substantiate the use of parametric tests.
METHODOLOGICAL CONSIDERATIONS AND FUTURE STUDIES
Every research study has room for improvement and this study is no exception. Admittedly, the estimation of PoR and p-values used here is rather liberal and may have some degree of Type I error attached to it. Thus, a replication study could estimate CVs for each normality test employed via quantiles as is traditionally done (although see section 2.2) and p-values could be combined via conservative methods such as the Stouffer method (see Vélez et al., under review). Also, other normality tests that are particularly robust to the distributions being studied could be considered. Equally important, other distributions that are a good fit for real data should be included in the simulation study. For instance, in the case of RT, data distributions such as Weibull and Log-Normal need to be studied. Another type of data commonly encountered in experimental research but less studied, is that of discrete n-point Likert-type data. Distributions that fit this type of data could be studied in the context of outliers and normality research as well. Yet, the studies suggested here should be preceded by research demonstrating how well various potential candidate distributions fit RT and Likert-type data (e.g., via AIC measures) and showing which distributions seem to give the best fits in both real and simulated RT and Likert-type data. Indeed, there should be research aimed at grounding the parameters of distributions fitting RT and Likert-type data into psychological processes of interest (e.g.,
Although some of the most commonly employed outlier elimination and transformation methods were addressed herein, other methods should also be studied. For instance, data truncation and outlier replacement are procedures also found in papers reporting experimental results in cognitive science (an example of data truncation can be found in
Finally, it can be contended that in principle, the procedures studied here should not only improve data’s normality but also their homogeneity. Thus, future studies should test the effects of the procedures studied here on the homogeneity of two or more batches of data. Canonical tests such as the Levene and the Brown-Forsythe test and recent robust versions of them (e.g.,
CONCLUSION
This paper sets out to offer an educated consensus on the recommended approach in cases where data need to be treated in order to submit to a parametric test. The results indicate that methods that transform the data in order to accommodate outliers lead to higher chances of normalization than methods that eliminate data points. Although some of the most commonly used elimination and transformation methods were studied herein, further methods need to be considered. Other distributions that can be used to model reaction time and Likert-type data should also be addressed in future studies.
Statements
Author contributions
Fernando Marmolejo-Ramos thanks Delphine Courvoisier and Firat Özdemir for their help with earlier versions of this project. Fernando Marmolejo-Ramos also thanks Kimihiro Noguchi for his help with his ‘nparLD’ R package and Xavier Romão for facilitating access to HPC facilities. Finally, the authors thank Robyn Groves and Rosie Gronthos for prooreading this manuscript.
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
Footnotes
1.^All of these tests are implemented in R in the following packages: “stats” package (SW test), “nortest” package (AD, KS, and SF tests), “lawstat” package (rJB test), and “normwhn.test” package (DH test). Note that the p-values computed for the SW test are based on the Royston method (Royston, 1982a,b, 1995), the version of the KS test in the “nortest” package is the Lilliefors, the critical values used to compute the rJB are obtained by an approximation to a χ distribution and with zero number of Monte Carlo simulations for the empirical critical values, and the p-value registered for the DH test was that labeled “sig.Ep” in the output.
2.^A pilot simulation showed that EGds of sample size 10 were in some cases trimmed down to seven when the elimination methods were used. Because such a low sample size affected the results of the AD and the DH tests, the sample size 10 was excluded from the final simulation study.
REFERENCES
1
AkbilgiçO.HoweA. J. (2011). A novel normality test using an identity transformation of the Gaussian function.Eur. J. Pure Appl. Math.4448–454.
2
Alizadeh NoughabiH.ArghamiN. R. (2011). Monte Carlo comparison of seven normality tests.J. Stat. Comput. Simul.81965–972. 10.1080/00949650903580047
3
AnsorgeU.KieferM.KhalidS.GrasslS.KönigP. (2010). Testing the theory of embodied cognition with subliminal words.Cognition116303–320. 10.1016/j.cognition.2010.05.010
4
BaayenR. H.MilinP. (2010). Analyzing reaction times.Int. J. Psychol. Res.312–28.
5
BartlettM. S. (1947). The use of transformations.Biometrics339–52. 10.2307/3001536
6
BeckmanR. J.CookR. D. (1983). Outliers. Technometrics25119–149.
7
BertelsJ.KolinskyR.MoraisJ. (2010). Emotional valence of spoken words influences the spatial orienting of attention.Acta Psychol.134264–278. 10.1016/j.actpsy.2010.02.008
8
BickelD. R.FrühwirthR. (2006). On a fast, robust estimator of the mode: comparisons to other robust estimators with applications.Comput. Stat. Data Anal.503500–3530. 10.1016/j.csda.2005.07.011
9
BlandJ. M.AltmanD. G. (1996). Transforming data.Br. Med. J.312:770. 10.1136/bmj.312.7033.770
10
BubD. N.MassonM. E. J. (2010). Grasping beer mugs: on the dynamics of alignment effects induced by handled objects.J. Exp. Psychol. Hum. Percept. Perform.36341–358. 10.1037/a0017606
11
CalkinsD. (1974). Some effects of non-normal distribution shape on the magnitude of the Pearson product moment correlation coefficient.Interam. J. Psychol.8261–287.
12
CousineauD. (2004). Merging race models and adaptive networks: a parallel race network.Psychon. Bull. Rev.11807–825. 10.3758/BF03196707
13
CousineauD.ChartierS. (2010). Outlier detection and treatment: a review.Int. J. Psychol. Res.358–67.
14
CowlesM.DavisC. (1982). On the origins of the 0.05 level of statistical significance.Am. Psychol.37553–558. 10.1037/0003-066X.37.5.553
15
De AlmeidaA.ElianS.NobreJ. S. (2008). Modificações e alternativas aos testes de Levene e de Brown e Forsythe para igualdade de variâncias e médias.Rev. Colomb. Estad.31241–260.
16
DondersF. C. (1969). On the speed of mental processes.Acta Psychol.30412–431. (Original published in 1869).
17
EngmannS.CousineauD. (2011). Comparing distributions: the two-sample Anderson-Darling test as an alternative to the Kolmogorov-Smirnoff test.J. Appl. Quant. Methods61–17.
18
HarriA.CobleK. H. (2011). Normality testing: two new tests using L-moments.J. Appl. Stat.381369–1379. 10.1080/02664763.2010.498508
19
HavasD. A.GlenbergA. M.RinckM. (2007). Emotion simulation during language comprehension.Psychon. Bull. Rev.14436–441. 10.3758/BF03194085
20
HeD.XuX. (2013). A goodness-of-fit testing approach for normality based on the posterior predictive distribution.Test221–18. 10.1007/s11749-012-0282-6
21
HeathcoteA.PopielS. J.MewhortD. J. K. (1991). Analysis of response time distributions: an example using the Stroop task.Psychol. Bull.109340–347. 10.1037/0033-2909.109.2.340
22
HockleyW. E. (1984). Analysis of response time distributions in the study of cognitive processes.J. Exp. Psychol. Learn. Mem. Cogn.10598–615. 10.1037/0278-7393.10.4.598
23
HohleR. (1965). Inferred components of reaction times as functions of foreperiod duration.J. Exp. Psychol.69382–386. 10.1037/h0021740
24
HopkinsG. W.KristoffersonA. B. (1980). Ultrastable stimulus-reponse latencies: acquisition and stimulus control.Percept. Psychophys.27241–250. 10.3758/BF03204261
25
HyndmanR. J.FanY. (1996). Sample quantiles in statistical packages.Am. Stat.50361–365.
26
JuddC. M.McClellandG. H.CulhaneS. E. (1995). Data analysis: continuing issues in the everyday analysis of psychological data.Annu. Rev. Psychol.46433–465. 10.1146/annurev.ps.46.020195.002245
27
LaBergeD. A. (1962). A recruitment theory of simple behavior.Psychometrika27375–396. 10.1007/BF02289645
28
LachaudC. M.RenaudO. (2011). A tutorial for analysing human reaction times: how to filter data, manage missing values, and choose a statistical model.Appl. Psychol.32389–416. 10.1017/S0142716410000457
29
LanceC. E.StewartA. M.CarrettaT. R. (1996). On the treatment of outliers in cognitive and psychomotor test data.Mil. Psychol.843–58. 10.1207/s15327876mp0801_4
30
LangloisD.CousineauD.ThiviergeJ.-P. (2014). Maximum likelihood estimators for truncated and censored power-law distributions show how neuronal avalanches may be misevaluated.Phys. Rev. E Stat Nonlin. Soft Matter Phys.8912709. 10.1103/PhysRevE.89.012709
31
LeysC.LeyC.KleinO.BernardP.LicataL. (2013). Detecting outliers: do not use standard deviation around the mean, use absolute deviation around the median.J. Exp. Soc. Psychol.49764–766. 10.1016/j.jesp.2013.03.013
32
MarkmanA. B.BrendlM. (2005). Constraining theories of embodied cognition.Psychol. Sci.166–10. 10.1111/j.0956-7976.2005.00772.x
33
Marmolejo-RamosF.González-BurgosJ. (2013). A power comparison of various tests of univariate normality on Ex-Gaussian distributions.Methodology9137–149.
34
Marmolejo-RamosF.MatsunagaM. (2009). Getting the most from your curves: exploring and reporting data using informative graphical techniques.Tutor. Quant. Methods Psychol.540–50.
35
McAuleyT.YapM.ChristS. E.WhiteD. A. (2006). Revisiting inhibitory control across the life span: insights from the ex-Gaussian distribution.Dev. Neuropsychol.29447–458. 10.1207/s15326942dn2903_4
36
McClellandJ. L. (1979). On the time relations of mental processes: a framework for analyzing processes in cascade.Psychol. Rev.86287–330. 10.1037/0033-295X.86.4.287
37
McGillW. J. (1963). “Stochastic latency mechanisms,” inHandbook of mathematical psychology,edsLuceR. D.BuschR. R.GalanterE. (New York: John Wiley and Sons), 309–360.
38
MillerJ. (1989). A warning about median reaction time.J. Exp. Psychol. Hum. Percept. Perform.14539–543. 10.1037/0096-1523.14.3.539
39
MillerJ. O.UlrichR. (2003). Simple reaction time and statistical facilitation: a parallel grains model.Cogn. Psychol.46101–115. 10.1016/S0010-0285(02)00517-0
40
MoranD. W.SchwartzA. B. (1999). Motor cortical representation of speed and direction during reaching.J. Neurophysiol.822676–2692.
41
MossH. E.McCormickS.TylerL. K. (1997). The time course of activation of semantic information during spoken word recognition.Lang. Cogn. Process.12695–731. 10.1080/016909697386664
42
MouriH. (2013). Log-normal distribution from a process that is not multiplicative but is additive.Phys. Rev. E88:042124. 10.1103/PhysRevE.88.042124
43
NoguchiK.GelY. R.BrunnerE.KonietschkeF. (2012). nparLD: an R software package for the nonparametric analysis of longitudinal data in factorial experiments.J. Stat. Softw.501–23.
44
OlivierJ.NorbergM. M. (2010). Positively skewed data: revisiting the Box-Cox power transformation.Int. J. Psychol. Res.368–75.
45
OrrJ. M.SackettP. R.DuBoisC. L. Z. (1991). Outlier detection and treatment in I/O psychology: a survey of researcher beliefs and an empirical illustration.Pers. Psychol.44473–486. 10.1111/j.1744-6570.1991.tb02401.x
46
OsborneJ. (2002). Notes on the Use of Data Transformation. Practical Assessment, Research and Evaluation.Available at: http://pareonline.net/getvn.asp?v=8&n=6 [accessed May 18 2009].
47
OtteE.HabelU.Schulte-RütherM.KonradK.KochI. (2011). Interference in simultaneously perceiving and producing facial expressions – evidence from electromyography.Neuropsychologia49124–130. 10.1016/j.neuropsychologia.2010.11.005
48
PereaM. (1999). Tiempos de reacción y psicología cognitiva: dos procedimientos para evitar el sesgo debido al tama no muestral.Psicológica2013–21.
49
PikeR. (1973). Response latency models for signal detection.Psychol. Rev.8053–68. 10.1037/h0033871
50
PylyshynZ. W.StormR. W. (1988). Tracking multiple independent targets: evidence for a parallel tracking mechanism.Spat. Vis.3179–197. 10.1163/156856888X00122
51
RaabD. H. (1962). Effects of stimulus-duration on auditory reaction-time.J. Am. Psychol.75298–301. 10.2307/1419616
52
RashidM. M.McKeanJ. W.KlokeJ. D. (2013). Review of rank-based procedures for multicenter clinical trials.J. Biopharm. Stat.231207–1227. 10.1080/10543406.2013.834919
53
RatcliffR. (1993). Methods for dealing with reaction time outliers.Psychol. Bull.114510–532. 10.1037/0033-2909.114.3.510
54
RatcliffR.MurdockB. B. (1976). Retrieval processes in recognition memory.Psychol. Rev.86190–214. 10.1037/0033-295X.83.3.190
55
RazaliN. M.WahY. B. (2010). “Power comparison of some selected normality tests,” inProceedings of the Regional Conference on Statistical Sciences, (RCSS’10),Kota Bharu, 126–138.
56
ReinJ. R.MarkmanA. B. (2010). Assessing the concreteness of relational representation.J. Exp. Psychol. Learn. Mem. Cogn.361452–1465. 10.1037/a0021040
57
RomãoX.DelgadoR.CostaA. (2010). An empirical power comparison of univariate goodness-of-fit tests for normality.J. Stat. Comput. Simul.80545–591. 10.1080/00949650902740824
58
RosenbergJ. L.GaskoM. (1983). “Comparing location estimators: trimmed means, medians, and trimean,” inUnderstanding Robust and Exploratory Data Analysis,edsHoaglinD.MostellerF.TukeyJ. (New York, NY: Wiley), 297–336.
59
RouderJ. N. (2000). Assessing the roles of change discrimination and luminance integration: evidence for hybrid race model of perceptual decision making in luminance discrimination.J. Exp. Psychol. Hum. Percept. Perform.26359–368. 10.1037/0096-1523.26.1.359
60
RousseeuwP. J.van DriessenK. (1999). A fast algorithm for the minimum covariance determinant estimator.Technometrics41212–223. 10.1080/00401706.1999.10485670
61
RoystonP. (1982a). Algorithm AS 181: the W test for normality.Appl. Stat.31176–180. 10.2307/2347986
62
RoystonP. (1982b). An extension of Shapiro and Wilk’s W test for normality to large samples.Appl. Stat.31115–124. 10.2307/2347973
63
RoystonP. (1995). Remark AS R94: a remark on algorithm AS 181: the W test for normality.Appl. Stat.44547–551. 10.2307/2986146
64
SchmiderE.ZieglerM.DanayE.BeyerL.BühnerM. (2010). Is it really robust? Reinvestigating the robustness of ANOVA against violations of the normal distribution assumption.Methodology6147–151.
65
SchwarzW. (2001). The ex-wald distribution as a descriptive model of response times.Behav. Res. Methods Instrum. Comput.33457–469. 10.3758/BF03195403
66
ThompsonG. L. (2006). An SPSS implementation of the non-recursive outlier deletion procedure with shifting z-score criterion (Van Selst and Jolicoeur, 1994).Behav. Res. Methods38344–352. 10.3758/BRM.38.2.344
67
UedaT. (2009). A simple method for the detection of outliers (trans. F. Marmolejo-Ramos and S. Kinoshita).Electron. J. Appl. Stat. Anal.267–76. (Original work published in 1996).
68
UlrichR.MaienbornC. (2010). Left-right coding of past and future in language: the mental timeline during sentence processing.Cognition117126–138. 10.1016/j.cognition.2010.08.001
69
UlrichR.MillerJ. (1993). Information processing models generating lognormally distributed reaction times.J. Math. Psychol.37513–525. 10.1006/jmps.1993.1032
70
van der LooM. P. J. (2010). Distribution Based Outlier Detection for Univariate Data.The Hague: Statistics Netherlands.
71
Van SelstM.JolicoeurP. (1994). A solution to the effect of sample size on outlier elimination.Q. J. Exp. Psychol. 47A, 631–650. 10.1080/14640749408401131
72
VélezJ. I.CorreaJ. C. (2014). Should we think of a different median estimator?Comun. Estadística711–17.
73
WestJ.ShlesingerM. (1990). The noise in natural phenomena.Am. Sci.7840–45.
74
WhelanR. (2008). Effective analysis of reaction time data.Psychol. Rec.58475–482.
75
WilcoxR. (2012). Introduction to Robust Estimation and Hypothesis Testing.Amsterdam: Elsevier.
76
WilcoxR. R. (1998). How many discoveries have been lost by ignoring modern statistical methods?Am. Psychol.53300–314. 10.1037/0003-066X.53.3.300
77
WilcoxR. R.KeselmanH. J. (2003). Modern robust data analysis methods: measures of central tendency.Psychol. Methods8254–274. 10.1037/1082-989X.8.3.254
78
YapB. W.SimC. H. (2011). Comparison of various types of normality tests.J. Stat. Comput. Simul.812141–2155. 10.1080/00949655.2010.520163
Summary
Keywords
Ex-Gaussian, reaction times, normality tests, outliers
Citation
Marmolejo-Ramos F, Cousineau D, Benites L and Maehara R (2015) On the efficacy of procedures to normalize Ex-Gaussian distributions. Front. Psychol. 5:1548. doi: 10.3389/fpsyg.2014.01548
Received
29 July 2014
Accepted
14 December 2014
Published
07 January 2015
Volume
5 - 2014
Edited by
Holmes Finch, Ball State University, USA
Reviewed by
Jocelyn Holden Bolin, Ball State University, USA; Julianne M. Edwards, Ball State University, USA
Copyright
© 2015 Marmolejo-Ramos, Cousineau, Benites and Maehara.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) or licensor are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
*Correspondence: Fernando Marmolejo-Ramos, Gösta Ekman Laboratory, Department of Psychology, Stockholm University, Frescati Hagväg 9A, Stockholm, SE-106 91, Sweden e-mail: fernando.marmolejo.ramos@psychology.su.se; Website: http://sites.google.com/site/fernandomarmolejoramos/
This article was submitted to Quantitative Psychology and Measurement, a section of the journal Frontiers in Psychology.
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.