ORIGINAL RESEARCH article

Front. Appl. Math. Stat., 10 May 2024

Sec. Statistics and Probability

Volume 10 - 2024 | https://doi.org/10.3389/fams.2024.1399837

Beta transformation of the Exponential-Gaussian distribution with its properties and applications

  • Department of Mathematics, Kotebe University of Education, Addis Ababa, Ethiopia

Abstract

This study introduces a five-parameter continuous probability model named the Beta-Exponential-Gaussian distribution by extending the three-parameter Exponential-Gaussian distribution with the beta transformation method. The basic properties of the new distribution, including reliability measure, hazard function, survival function, moment, skewness, kurtosis, order statistics, and asymptotic behavior, are established. Using the acceptance-rejection algorithm, simulation studies are conducted. The new model is fitted to the simulated and real data sets, and its performance is reported. The Beta-Exponential-Gaussian distribution is found to be more flexible and has better performance in many aspects. It is suggested that the new distribution would be used in modeling data having skewness and bimodal distribution.

1 Introduction

The beta distribution family gained popularity a few years ago, and various studies have been conducted on beta compounding with other distributions such as Beta-Frechet [], Beta-Gumbel [], Beta-Pareto [], Beta-Gamma [], Beta-Gompertz [], Beta-Normal[], Beta-Power [], Beta-Weibull [], Beta-Exponentiated Weibull [], Beta-Modified Weibull [], Beta-Extended Weibull[], the Beta-Generalized Logistic [], Beta-Log-Logistic [], Beta-Exponential [], Beta-Generalized Exponential [], Beta-Moyal [], Beta-Generalized Half-Normal [], Beta-Dagum [], Beta-Laplace [], Beta-Burr XII [], Beta-Generalized Pareto [], Beta-Cauchy [], Beta-Half-Cauchy [], Beta-Exponentiated Pareto [], Beta-Log-Logistic [], Beta-Type I Generalized Half Logistic [], Beta-Bessel Distributions [], Beta-Rayleigh Distribution [], Beta-Exponentiated Lindley [], Beta-Lindley [], Beta-Lindley-Poisson Distribution [], and Beta-Lindley Geometric Distribution [].

This study presents a five-parameter distribution called the Beta-Exponential-Gaussian distribution, which extends the Ex-Gaussian distribution by adding two additional parameters to control its skewness and kurtosis.

In cognitive psychology research, the Ex-Gaussian distribution is used to examine the semantic stroop effect [], reading eye movements [], and response time distributions []. Additionally, it is utilized to mimic fixation lengths [], which are frequently employed in eye-tracking research as a gauge of cognitive processes.

This investigation unfolds to examine the new Beta-Ex-Gaussian distribution. In Section 2, a parent sub-model, the basic Ex-Gaussian distribution is introduced. The transformation technique for the proposed distribution is explained in Section 3. Methodically, Section 4 examines the Beta-Ex-Gaussian distribution. Advancing to Section 5, visual representations of the probability density function (PDF) and cumulative distribution function (CDF) of the Beta-Ex-Gaussian distribution are presented. Section 6 investigates statistical distinctions by deriving expressions for moments. The discourse within Section 7 explores the sphere of order statistics. Survival and hazard functions are scrutinized in Section 8. In Section 9, attention is directed toward estimation methodologies, with an emphasis on maximum likelihood. Section 10 offers simulation results from a comprehensive acceptance-rejection process. Section 11 expounds on four practical applications. Finally, Section 12 provides a conclusive conclusion, thereby concluding the article.

2 Basic exponential-Gaussian distribution

The exponential-Gaussian (Ex-Gaussian), also called the exponentially modified Gaussian (EMG) distribution, is a probability distribution that convolutions exponential and normal random variables and is used in signal processing, finance, and neuroscience for skewness and heavy tails data. It has a closed-form probability density function and a cumulative distribution function, making it useful for statistical analysis []. For ∀μ ∈ ℝ, σ, λ > 0, and ∀x ∈ ℝ, its PDF, CDF, Hazard, and Survival functions are given in the Equations (14), respectively.

where erfc is the complementary error function defined by .

3 Transformation techniques

Many transformation methods were applied to generate new probability distributions. This study uses the beta-generated (B-G) approach to generate continuous probability distributions. This technique modifies a base distribution using a generator function, resulting in a variety of shapes and characteristics. The beta distribution is used as the generator, adding more parameters to fit different shapes and determining skewness by the sum of all forms [].

Let f(x; Θ) and F(x; Θ) be the probability density function and the cumulative distribution function (cdf) of a random variable X, respectively, where Θ is a p × 1 parameter vector, and then the cumulative distribution function is generated by applying [, ]

where α > 0 and β > 0 are two additional parameters whose role is to introduce skewness and vary the tail weight. is the beta function, is the incomplete beta function with B(α, β) = B1(α, β), and is the incomplete beta function ratio. The corresponding probability density function of Equation (5) is then given as follows:

This family of distributions can be considered as generalization of the distribution of order statistics [] for the random variable X with CDF F(x). When α and β are integers, Equation (6) is the αth order statistic of the random sample of size (α+β−1).

4 The new Beta-Ex-Gaussian(BExG) distribution

Let F(x; Θ) be the baseline CDF of an Ex-Gaussian, a continuous random variable, with Θ = (μ, σ, λ) as parameter vectors. We now introduce the five-parameter Beta-Ex-Gaussian (BExG) distribution by taking G(x) from Equation (5), the CDF of Equation (6). Substituting F(x; Θ) in Equation (5) by the CDF Equation (2) yields the CDF of the BExG as follows:

where α > 0 and β > 0 are new parameters that controls the skewness and kurtosis of the distribution, i.e., the distribution's shape, and ∀μ ∈ ℝ, σ, λ > 0, and ∀x ∈ ℝ are the location, scale, and rate parameters, respectively. A random variable X with the CDF Equation (7) is said to have a BExG distribution and will be denoted by X~BExG(Φ), where Φ = (μ, σ, λ, α, β).

For any values of α and β, we can write Equation (5) in terms of the well-known hypergeometric function (see [, , ]) given as follows:

Theorem 1. The new random variable X for x ∈ ℝ, λ > 0, μ ∈ ℝ, σ > 0, α > 0, and β > 0 having cumulative distribution function (CDF) and probability density function (PDF) expressed as follows:

  • Cumulative distribution function (CDF):

  • Probability density function (PDF):

    g(x)≥0

where and (a)k = a(a+1)⋯(a+k−1).

Proof.

This function, g(x) in (2a), is always non-negative, which is trivial to prove, and

To prove (2b) integrates is 1 over the real number, we used Equation (6) for g(x)

Use integration by substitution: u = F(x) and differentiate u with respect to x, yielding:

The integral is a special case of the beta function, defined as:

Table 1 shows that some closed-form expansions for Equation (5) distribution based on real, non-integer, or integer parameters.

Table 1

No.Cumulative distribution functionParameterSource
1.For any values of α and β[, ]
2.For real non integer β[]
3.For integer β[]
4.For integer α[]

Some closed-form expansions for Equation (5) distribution based on real, non-integer, or integer parameters.

Now, by changing variables, let we get . To change limit of integration from 0 to ∞, consider t → 0, u → 0;t → 1, u → ∞.

Therefore, we have:

Therefore, we have shown that

Corollary 1.1. Asymptotic properties:

  • i.

  • ii.

Proof. To prove (i), the property for the confluent hypergeometric function 2F1 from the book [] is:

where Γ(·) is the Gamma function and the properties of the beta function, respectively.

Then,

We have demonstrated conclusively that:

To prove (ii)

Therefore, .

4.1 Special cases

The following are some special cases of the BExG distribution that we examine:

  • When both α = β = 1, the BExG distribution Equation (10) simplifies to the Ex-Gaussian distribution Equation (1). This distribution is characterized by parameters μ, σ, and λ as derived from the study mentioned in Golubev [].

  • In the scenario where α = β = 1 and λ → ∞, the BExG distribution Equation (10) transforms into the Gaussian distribution. This transformation is described by the parameters μ and σ proposed by the study mentioned in Marmolejo-Ramos et al. [].

  • When β = 1, the BExG distribution Equation (3.2) reduces to the generalized exponential-Gaussian distribution. This reduction is governed by parameters λ, μ, σ, and α according to the study mentioned in Marmolejo-Ramos et al. [].

  • In the case where β = 1, λ → ∞, and α≠1, the BExG distribution Equation (6) conforms to the power-normal distribution. This distribution is characterized by parameters μ, σ, and α, as studied by Chen et al. [] and Marmolejo-Ramos et al. [].

5 Plots of the probability distribution

In this section, the BExG probability density function (PDF) and cumulative distribution function (CDF) plots in the Figures 15 are presented, showcasing various selected parameter values.

Figure 1

Figure 2

Figure 3

Figure 4

Figure 5

A bimodal PDF (Figure 5) shows two different peaks or modes, indicating the possibility of two unique processes or sub-populations with various traits. The relative heights of each peak indicate the frequencies and separations between them, while each peak itself represents a mode or cluster within the data.

Some of the shape properties of the uni-modal Beta-Ex-Gaussian distribution include:

  • The proposed distribution is a right-skewed distribution when α > β. As α increases with β fixed, the degree of right skewness increases.

  • Conversely, the distribution demonstrates left-skewness when α < β. As β increases with α fixed, the degree of left skewness increases.

  • Notably, when α≠β, the distribution demonstrates skewness.

  • In the case where α = β, the distribution attains symmetry.

  • When both α > 1 and β > 1, the distribution is characterized as leptokurtic, exhibiting a higher peak.

  • Conversely, when both α < 1 and β < 1, the distribution assumes a platykurtic form, featuring a heavier tail.

  • Specifically, when α = 1, β = 1, and λ → ∞, the distribution is categorized as mesokurtic, closely resembling a normal distribution.

In general, the new distribution provides more flexible and versatile shapes that are different from those of normal and exponential-normal distributions.

6 Moments of the Beta-Ex-Gaussian

In probability and statistics, expectations of powers are moments of random variables, where the first is the expectation and the second is the variance, or the second central moment [].

For β ∈ ℝ\ℤ, we have the power series

If β ∈ ℤ, the index j in the sum stops at β−1 [].

For α ∈ ℝ\ℤ, F(x)α+j−1 can be expanded:

Theorem 2. When α and β are integers, the nth moment of the Beta-Ex-Gaussian random variable BExG(α, β, μ, σ, λ) is given in the Equation (13) as follows:

where

Proof. The book [] explains that for a random variable X in Ln, the nth moment, the mean, and the central moment respectively are:

Using Equations (11, 12), and Karr [] we can prove it as follows:

where

The variance, skewness, and kurtosis measures can be calculated and related using the study mentioned in Jafari et al. [] and Bury [] in the Equations (1416) respectively.

The skewness and kurtosis measures are controlled mainly by the parameters α and β, and Figure 6 illustrates their variation using μ = 0, σ = 1, and λ = 1.

Figure 6

The skewness and kurtosis measures are controlled mainly by the parameters α and β, and Figures 610 illustrates their variation using various parametric values for μ, σ, and λ. More figures are illustrated in Appendix A. The skewness and kurtosis demonstrate strictly increasing, strictly decreasing, and U-shaped behaviors, which are interesting in some applications of the model.

Figure 7

Figure 8

Figure 9

Figure 10

Moreover, Tables 2, 3 illustrate how the mean and median values of a distribution offer insights into its central tendency. Skewness gauges the asymmetry of the distribution, with positive values suggesting longer or fatter right tails and negative values indicating the opposite. Kurtosis, on the other hand, measures the tailedness of the distribution, with higher values indicating heavier tails or more outliers. Additionally, the impact of α on the distribution's skewness and kurtosis can be observed; a higher α value may result in a more asymmetric distribution.

Table 2

βαMeanMedianStd DevVarianceSkewnessKurtosis
30.2900.1470.4760.2261.9956.462
40.1760.0840.3490.1222.4729.399
510.1120.0510.2640.0703.00813.142
60.0730.0330.2040.0423.60618.080
100.0160.0070.0810.0076.87060.531
30.6330.0630.6200.3841.007–2.719
40.4330.0330.4970.2471.2260.791
520.3050.0190.4070.1661.4913.119
60.2190.0110.3370.1131.7885.061
100.0640.0020.1640.0273.29415.304
30.9340.0300.6620.4390.637–20.380
40.6880.0140.5590.3120.721–10.548
530.5200.0070.4820.2320.865–4.784
60.3980.0040.4190.1751.042–1.255
100.1460.0010.2420.0581.9586.057
31.1800.0160.6660.4440.497–55.500
40.9120.0070.5760.3320.488–34.682
540.7220.0030.5110.2610.547–20.967
60.5790.0020.4580.2100.646–12.291
100.2510.0000.3020.0911.2531.265
32.0180.0020.6300.3960.534–629.522
41.7230.0000.5440.2950.408–602.938
5101.5070.0000.4890.2400.326–535.531
61.3360.0000.4520.2040.270–456.325
100.8850.0000.3670.1350.178–199.409

The mean, variance, median, standard deviation, skewness, and kurtosis for μ = 0, λ = 1, and σ = 1.

Table 3

βαMeanMedianStd DevVarianceSkewnessKurtosis
132.2070.3291.2851.6511.072–46.946
42.4980.2491.2781.6331.092–82.296
52.7230.2001.2731.6191.114–120.255
62.9050.1671.2691.6111.131–158.980
103.4100.1001.2661.6031.160–310.309
231.3340.0800.8310.6910.683–36.037
41.6020.0490.8250.6810.640–81.185
51.8130.0330.8160.6660.651–142.007
61.9860.0240.8080.6540.675–214.260
102.4710.0090.7940.6300.744–559.040
330.9340.0300.6620.4390.637–20.380
41.1800.0160.6660.4440.497–55.500
51.3800.0090.6590.4350.460–111.550
61.5470.0060.6510.4240.464–187.607
102.0180.0020.6300.3960.534–629.522
430.6880.0140.5590.3120.721–10.548
40.9120.0070.5760.3320.488–34.682
51.1020.0030.5750.3300.391–78.069
61.2630.0020.5680.3220.362–143.753
101.7230.0000.5440.2950.408–602.938
1030.1460.0010.2420.0581.9586.057
40.2510.0000.3020.0911.2531.265
50.3680.0000.3410.1160.821–4.991
60.4860.0000.3620.1310.543–16.548
100.8850.0000.3670.1350.178–199.409

The mean, variance, median, standard deviation, skewness, and kurtosis for μ = 0, λ = 1, and σ = 1.

7 Order statistics

The density of the ith order statistic Xi:n, gi:n(x) say, in a random sample of size n from the BExG distribution is obtained from the well-known formula [] is given in the Equation (17) as follows.

Or let X1, X2, ..., Xn be a random sample of size n from BExG(μ, σ2, λ, α, β). Then the pdf and cdf of the ith order statistic, say Xi:n, are given by Jafari et al. []

respectively, where . Here, we use an equation by Gradshteyn and Ryzhik [] and Jafari et al. [], for a power series raised to a positive integer n

where the coefficients cn, r (for r = 1, 2, ...) are easily determined from the recurrence Equation (20)

where . The coefficient cn, r can be calculated from cn, 0, ..., cn, r−1 and hence from the quantities b0, ..., br.

The Equations (18, 19) can be written as follows:

Therefore, the sth moment of Xi:n is as follows

which cannot be an explicit expression.

8 Survival and hazard functions

In this section, we refer to Klein and Moeschberger [] to underscore the fundamental significance of the likelihood of an individual surviving beyond time x as a key metric in characterizing time-to-event events. The survival function, denoted by the random variable X, is described in the Equation (21) as follows (see Klein and Moeschberger [] for details).

Furthermore, the hazard function, also referred to as the conditional failure rate, is a critical parameter in survival analysis with broad applications in fields such as economics, epidemiology, demography, and stochastic processes. For a continuous random variable X, the hazard function is defined as follows:

Theorem 3. The survival function and hazard function for the new random variable are as follows:

  • Survival function:

  • Hazard function:

Proof. Equations (21, 22) can be used to easily obtain the proof for Equations (23, 24), respectively.

Corollary 3.1. Asymptotic behaviors:

Proof.

Therefore, .

Here, we establish a series of plots to visually represent the hazards and survival functions of the BExG distribution. This distribution is characterized by varying parameter values, including α, β, μ, σ, and λ. Through these plots, we aim to provide insights into the probability density and hazards associated with the random variable under consideration.

Based on parameter values, the hazard function can take on a variety of shapes, including monotonically increasing, decreasing, bimodal, parabolas, and bumping shapes. The hazard rate function appears to converge to a fixed value over a large range of x. Figures 1116 reveal that larger α and β values result in higher hazard rates compared with those of the base distribution. Overall, the new distribution manifests shapes and patterns that are distinct from the base distribution. Specifically, the shape parameters play a crucial role in generating the interesting shapes observed in the new BExG probability distribution. The instantaneous rate of an event of interest is described by a bimodal hazard function, where each peak denotes a higher risk period.

Figure 11

Figure 12

Figure 13

Figure 14

Figure 15

Figure 16

Now, we go into the analysis of specific hazard function plots (Figure 16) derived from the BExG distribution across a range of parameter values, with particular attention to the scale (σ) and shape (α) parameters. Our aim is to elucidate their influence on the hazard profiles of the distribution and their implications for practical applications.

9 Parameter estimation

Let X1, X2, ⋯ , Xn be n i.i.d. random observations generated from the new distribution g(x; Φ).

The likelihood and log-likelihood function for the new BExG distribution are given by Equations (25, 26), respectively.

Here, Φ = {α, β, μ, σ, λ}.

The Maximum Likelihood Estimate (MLE), denoted as or , is the set of parameter values, which maximizes the likelihood function or, equivalently, the log-likelihood function []

Let Φ = {α, β, μ, σ, λ} and the log-likelihood function be denoted as ℓ(Φ|x).

The Maximum Likelihood Estimator (MLE) for each parameter is obtained by taking partial derivatives of the log-likelihood function and setting the equation equal to zero. i.e.,

Since it was not easy to determine the estimated parameters analytically, we used numerical optimization approaches.

10 Simulations

This section explores the use of the acceptance-rejection algorithm, a fundamental method in statistical simulation used to produce random samples from probability distributions that are difficult to sample directly. In particular, we study the Beta-Ex-Gaussian (BExG) distribution, which is a probabilistic complex that combines elements of the beta, exponential, and Gaussian distributions. The most important steps in the acceptance-rejection algorithm by the study mentioned in Robert and Casella [] are as follows:

  • Step 1. Generate a random variable Y from a proposal density function g(y).

  • Step 2. Generate a uniform random variable, U, which is independent of Y.

  • Step 3. If , accept the proposed Y and set X = Y; otherwise, go to step 1.

where g(y) is a known distribution close to f(y), and M is a constant number that is the upper bound such that . In our case, f(y) is the new Beta-Ex-Gaussian density function.

The simulation involves generating N = 10,000 samples from the target distribution. Subsequently, the algorithm visually presents the simulated samples through density plots and histograms to illustrate their distribution, as illustrated in Figures 1719. The histograms show that the distribution can be skewed distribution depending on the parameter values.

Figure 17

Figure 18

Figure 19

To assess the effectiveness of Maximum Likelihood Estimators (MLEs) for the parameters of the Beta-Ex-Gaussian (BExG) distribution, we conduct simulations across varying sample sizes. The evaluation encompasses three distinct cases characterized by the true parameter sets denoted as Φtrue = (α, β, μ, σ, λ):

  • Case I: α = 2, β = 2, μ = 0, σ = 1, λ = 1,

  • Case II: α = 5, β = 2, μ = 0, σ = 1, λ = 1, and

  • Case III: α = 2, β = 6, μ = 0, σ = 1, λ = 1

The MLE, for each parameter can be evaluated using two accuracy measures: the bias and the mean square error (MSE).

Table 4 shows biases, MSE, and MLE for simulated data. Larger sample sizes indicate closer convergence toward true parameter values, while lower MSE values indicate better estimation accuracy. Bias represents systematic estimation method overestimation or underestimation, while MSE measures estimator accuracy. In general, from Table 4, we find that the randomness of the sample size affects the fluctuation estimation accuracy, MSE, and bias of model parameters.

Table 4

Case ICase IICase III
nParams.MLEMSEBiasMLEMSEBiasMLEMSEBias
100α1.9970.00001–0.0034.5570.196–0.4431.9140.0074–0.086
β1.5350.2164–0.4651.5680.1868–0.4325.4980.2521–0.502
μ0.5020.25200.5020.0860.00740.0860.0610.00370.061
σ1.0400.00160.0401.6990.48810.6990.5820.1746–0.418
λ0.0100.9799–0.9900.0080.9832–0.9920.7920.0432–0.208
200α1.7290.1879–0.4334.5670.1938–0.4401.9170.0068–0.083
β1.6280.1893–0.4351.5650.2877–0.5365.4970.2526–0.503
μ–0.3290.01330.1150.1150.0340.1860.0610.00380.061
σ1.1460.51720.7191.7190.104–0.3230.5840.1734–0.416
λ0.0240.9887–0.9940.0060.8055–0.8970.7880.0452–0.212
500α1.8580.0201–0.144.5710.184–0.4291.9100.0081–0.09
β1.5960.1636–0.4041.5600.1936–0.445.4980.2515–0.502
μ0.2230.04990.2230.1340.0180.1340.0360.00130.036
σ1.0080.00010.0081.7390.54590.7390.5910.1675–0.409
λ0.1050.8009–0.8950.0090.981–0.9910.7930.043–0.207
1,000α1.8640.0184–0.1364.5690.1853–0.4311.9140.0073–0.086
β1.58810.1695–0.4121.5530.1998–0.4475.4980.2523–0.502
μ0.1630.026610.1630.2520.06360.2520.0410.00170.041
σ1.0670.004470.0670.8290.0291–0.1710.5920.1667–0.408
λ0.0960.81689–0.9040.1600.7049–0.840.7860.0457–0.214

Maximum likelihood estimates (MLE), mean squared errors (MSE), and bias of simulated data across different sample sizes and cases.

In our study, a simulated dataset to evaluate the goodness-of-fit of the Beta-Ex-Gaussian distribution using various metrics. The results showed that the new proposed distribution fits better than its base Ex-Gaussian distribution, as shown in Table 5 and Figure 20.

Table 5

ModelMLEWAValueAICBICCAICHQIC
Ex-G0.1000.976336.695679.391692.035679.439684.353
BExG0.0840.835243.178496.3562517.429496.478504.625

The ML estimates, W, A, log-likelihood, BIC, AIC, and CAIC for simulated data.

Figure 20

11 Applications

In this part, we analyze four actual data sets to demonstrate that the Beta-Ex-Gaussian (BExG) distribution fits better than the Ex-Gaussian (Ex-G) distribution.

To test, measures of goodness-of-fit can be applied in comparison to some other models. Mainly, we use Log-likelihood, statistic Cramer-von misses (W), statistic Anderson Darling (A), Akaike information criterion (AIC), the corrected Akaike information criterion (CAIC), Bayesian information criterion (BIC), and HannanQuinn information criterion (HQIC).

Data set 1: plasma concentrations-Indometh data

The Indometh data, a set of pharmacokinetic data from plasma concentrations of the indometacin vector, is an R-built-in data frame used in our study as data set 1 and is given as follows:

1.5, 0.94, 0.78, 0.48, 0.37, 0.19, 0.12, 0.11, 0.08, 0.07, 0.05, 2.03, 1.63, 0.71, 0.7, 0.64, 0.36, 0.32, 0.2, 0.25, 0.12, 0.08, 2.72, 1.49, 1.16, 0.8, 0.8, 0.39, 0.22, 0.12, 0.11, 0.08, 0.08, 1.85, 1.39, 1.02, 0.89, 0.59, 0.4, 0.16, 0.11, 0.1, 0.07, 0.07, 2.05, 1.04, 0.81, 0.39, 0.3, 0.23, 0.13, 0.11, 0.08, 0.1, 0.06, 2.31, 1.44, 1.03, 0.84, 0.64, 0.42, 0.24, 0.17, 0.13, 0.1, and 0.09.

Data set 2: carbon fiber data

Data set 2 consists of observations on the breaking stress of carbon fibers. The data set that was studied was used by the study mentioned in Kuttan Pillai et al. [] to study a new generalization of Pareto distribution and its applications. We fit to this data into the new model, BExG and Ex-Gaussian, and compare the results.

3.70 2.74 2.73 2.50 3.60 3.11 3.27 2.87 1.47 3.11 4.42 2.41 3.19 3.22 1.69 3.28 3.09 1.87 3.15 4.90 3.75 2.43 2.95 2.97 3.39 2.96 2.53 2.67 2.93 3.22 3.39 2.81 4.20 3.33 2.55 3.31 3.31 2.85 2.56 3.56 3.15 2.35 2.55 2.59 2.38 2.81 2.77 2.17 2.83 1.92 1.41 3.68 2.97 1.36 0.98 2.76 4.91 3.68 1.84 1.59 3.19 1.57 0.81 5.56 1.73 1.59 2.00 1.22 1.12 1.71 2.17 1.17 5.08 2.48 1.18 3.51 2.17 1.69 1.25 4.38 1.84 0.39 3.68 2.48 0.85 1.61 2.79 4.70 2.03 1.80 1.57 1.08 2.03 1.61 2.12 1.89 2.88 2.82 2.05 3.65.

Data set 3: survival times (in days) of 72 guinea pigs

Data on the survival times (in days) of 72 guinea pigs infected with virulent tubercle bacilli, observed, reported, and used by the study mentioned in Mukherjee et al. [] to study estimators of the PDF and CDF of the one-parameter polynomial exponential distribution.

0.1, 0.33, 0.44, 0.56, 0.59, 0.72, 0.74, 0.77, 0.92, 0.93, 0.96, 1, 1, 1.02, 1.05, 1.07, 1.07, 1.08, 1.08, 1.08, 1.09, 1.12, 1.13, 1.15, 1.16, 1.2, 1.21, 1.22, 1.22, 1.24, 1.3, 1.34, 1.36, 1.39, 1.44, 1.46, 1.53, 1.59, 1.6, 1.63, 1.63, 1.68, 1.71, 1.72, 1.76, 1.83, 1.95, 1.96, 1.97, 2.02, 2.13, 2.15, 2.16, 2.22, 2.3, 2.31, 2.4, 2.45, 2.51, 2.53, 2.54, 2.54, 2.78, 2.93, 3.27, 3.42, 3.47, 3.61, 4.02, 4.32, 4.58, and 5.55.

Data set 4: Kevlar 49/epoxy strand failure time data (pressure at 90%)

Life data set that displays the stress rupture life in hours of Kevlar 49/epoxy strands under continuous, sustained stress level pressure until failure. The data set has been previously used in the study mentioned in Al-Aqtash et al. [] to illustrate the usefulness of the Gumbel–Weibull distribution (GWD) when compared with the exponentiated-Weibull, beta normal, and generalized half normal. 0.01, 0.01, 0.02, 0.02, 0.02, 0.03, 0.03, 0.04, 0.05, 0.06, 0.07, 0.07, 0.08, 0.09, 0.09, 0.10, 0.10, 0.11, 0.11, 0.12, 0.13, 0.18, 0.19, 0.20, 0.23, 0.24, 0.24, 0.29, 0.34, 0.35, 0.36, 0.38, 0.40, 0.42, 0.43, 0.52, 0.54, 0.56, 0.60, 0.60, 0.63, 0.65, 0.67, 0.68, 0.72, 0.72, 0.72, 0.73, 0.79, 0.79, 0.80, 0.80, 0.83, 0.85, 0.90, 0.92, 0.95, 0.99, 1.00, 1.01, 1.02, 1.03, 1.05, 1.10, 1.10, 1.11, 1.15, 1.18, 1.20, 1.29, 1.31, 1.33, 1.34, 1.40, 1.43, 1.45, 1.50, 1.51, 1.52, 1.53, 1.54, 1.54, 1.55, 1.58, 1.60, 1.63, 1.64, 1.80, 1.80, 1.81, 2.02, 2.05, 2.14, 2.17, 2.33, 3.03, 3.03, 3.34, 4.20, 4.69, and 7.89.

The BExG distribution yields the smallest values of W, A, Log-likelihood function, AIC, BIC, CAIC, and HQIC statistics, as shown in Tables 69. Based on these statistics, it can be concluded that the BExG model outperforms the Ex-Gaussian distribution in fitting the data. The plots of the densities (alongside the data histogram) and cumulative distribution functions (with an empirical distribution function) are provided in Appendix B. These plots demonstrate also that the BExG model offers a superior fit compared with the Ex-Gaussian model.

Table 6

ModelMLEWAValueAICBICCAICHQIC
Ex-G0.5543.350131.318268.637275.206269.025271.233
0.615
BExG0.3262.08043.79697.592108.54198.592101.919

MLEs, W, A, log-likelihood, BIC, AIC, and CAIC for data set 1.

Table 7

ModelMLEWAValueAICBICCAICHQIC
Ex-G0.2171.142176.287358.574366.390358.824361.737
BExG0.1120.574147.257304.514317.540305.152309.786

MLEs, W, A, log-likelihood, BIC, AIC, and CAIC for data set 2.

Table 8

ModelMLEWAValueAICBICCAICHQIC
Ex-G0.0560.37396.177198.355205.185198.708201.074
BExG0.0820.46893.298196.597207.980197.506201.128

MLEs, W, A, log-likelihood, BIC, AIC, and CAIC for data set 3.

Table 9

ModelMLEWAValueAICBICCAICHQIC
Ex-G0.6814.570181.071368.142375.987368.389371.318
BExG0.1811.239113.605237.210250.285237.841242.503

MLEs, W, A, log-likelihood, BIC, AIC, and CAIC for data set 4.

12 Conclusion

The aim of this study is to develop a new five-parameter continuous probability distribution, named the Beta-Exponential-Gaussian (BExG) distribution using the method of beta generator. The BExG distribution includes Ex-Gaussian, generalized Ex-Gaussian, normal, and power-normal probability distributions as special cases. It is a new contribution to the Statistical and Probability theory. The research contributes to enhancing modeling capabilities of the base Exponential-Gaussian distribution by proposing the new Beta Exponential-Gaussian distribution which is found to be a promising alternative for data analysis in different application areas including survival and reliability. The basic properties of the new distribution, including reliability measure, hazard function, survival function, moment, skewness, kurtosis, order statistics, and asymptotic behavior, are established. The acceptance-rejection algorithm for simulation is presented. The new model is fitted to the simulated and real data sets, and its performance is revealed. The distribution can model data sets having a distributional nature of various skewness, kurtosis, heavier tails, uni-model, bi-modal, and asymmetric properties. The Beta-Exponential-Gaussian distribution is found to be more flexible, and so, we recommend it to be used for applications.

Statements

Data availability statement

The original contributions presented in the study are included in the article/Supplementary material, further inquiries can be directed to the corresponding author.

Author contributions

KT: Conceptualization, Data curation, Formal analysis, Investigation, Methodology, Writing – original draft, Writing – review & editing. AG: Formal analysis, Investigation, Validation, Supervision, Writing – review & editing.

Funding

The author(s) declare that no financial support was received for the research, authorship, and/or publication of this article.

Conflict of interest

The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.

Publisher’s note

All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article, or claim that may be made by its manufacturer, is not guaranteed or endorsed by the publisher.

Supplementary material

The Supplementary Material for this article can be found online at: https://www.frontiersin.org/articles/10.3389/fams.2024.1399837/full#supplementary-material

References

Summary

Keywords

Ex-Gaussian distribution, Beta-Ex-Gaussian distribution, beta-generator, acceptance-rejection algorithm, MLE

Citation

Tesfaw KW and Goshu AT (2024) Beta transformation of the Exponential-Gaussian distribution with its properties and applications. Front. Appl. Math. Stat. 10:1399837. doi: 10.3389/fams.2024.1399837

Received

12 March 2024

Accepted

15 April 2024

Published

10 May 2024

Volume

10 - 2024

Edited by

Mohammad G. M. Khan, University of the South Pacific, Fiji

Reviewed by

Nesar Ahmad, Tilka Manjhi Bhagalpur University, India

Fuxia Cheng, Illinois State University, United States

Updates

Copyright

*Correspondence: Kumlachew Wubale Tesfaw

Disclaimer

All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.

Outline

Figures

Cite article

Copy to clipboard


Export citation file


Share article

Article metrics