Next Article in Journal
A Sequential Markov Probabilistic Aggregation Algorithm for Causal Emergence in Markov Aggregation
Next Article in Special Issue
Bayesian Sampling with Approximate Transport Geometry via Residual-Slice Correction
Previous Article in Journal
On the Importance of Separation and Labelling on the Hypersphere
Previous Article in Special Issue
Bayesian Estimation of Marginal Quantiles with Missing Data in a Multivariate Regression Framework
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

A Computational Bayesian Framework for Entropy-Based Inference in the Inverse Gaussian Distribution Under Progressive Type-II Censoring

by
Mohamed A. T. El-Shahat
1,
Reman Abo Hashem
1,
Doaa Basalamah
2,
Tmader Alballa
3,* and
Wael S. Abu El Azm
1
1
Department of Statistics and Insurance, Faculty of Commerce, Zagazig University, Zagazig 44519, Egypt
2
Mathematics Department, Faculty of Science, Umm Al-Qura University, P.O. Box 24231, Makka 21955, Saudi Arabia
3
Department of Mathematical Sciences, College of Science, Princess Nourah bint Abdulrahman University, P.O. Box 84428, Riyadh 11671, Saudi Arabia
*
Author to whom correspondence should be addressed.
Entropy 2026, 28(8), 871; https://doi.org/10.3390/e28080871
Submission received: 4 June 2026 / Revised: 18 July 2026 / Accepted: 30 July 2026 / Published: 2 August 2026
(This article belongs to the Special Issue Advances in Bayesian Statistics)

Abstract

Estimating entropy measures under progressive censoring poses a significant challenge in reliability and lifetime analysis. Despite the widespread use of the inverse Gaussian distribution in modeling skewed lifetime data, a comprehensive inferential framework for its Shannon and Rényi entropies under progressive Type-II censoring remains absent. This paper develops a unified Bayesian framework integrating maximum likelihood and Bayesian inference under squared error, general entropy, and LINEX loss functions, employing Lindley’s approximation, importance sampling, and Markov chain Monte Carlo methods. Two parameter configurations were examined to assess robustness under varying likelihood surface complexity, along with prior sensitivity analysis. Results demonstrate that Markov chain Monte Carlo and importance sampling are the only consistently reliable methods across all scenarios, whereas maximum likelihood suffered severe bias and collapse of Wald interval coverage under challenging settings, and Lindley’s approximation exhibited numerical instability at small samples. Shannon entropy proved substantially more sensitive to parameter variation than Rényi entropy, with the LINEX and general entropy loss functions showing superior performance. The study offers clear practical guidance, validated on real lifetime data.

1. Introduction

Entropy is a fundamental measure of uncertainty in probability distributions, reflecting the average information content of a sample. Shannon entropy, introduced by Shannon [1], is the most widely used formulation of this concept, later generalized by Rényi through an additional parameter α that provides flexibility in tuning the measure’s sensitivity. Entropy is particularly valuable in reliability and lifetime data analysis, outperforming traditional measures such as the mean and variance in characterizing skewed and heavy-tailed distributions. Shannon entropy is defined as:    
H = E ln f ( x ) = 0 ln f ( x ) f ( x ) , d x
Rényi [2] extended this measure through an order parameter α , defined as:
H α = 1 1 α ln 0 f ( x ) α , d x , α ( 1 ) > 0
Entropy has been applied in diverse fields, including password security [3], agricultural economics [4], seismology [5], and software reliability [6]. Shannon entropy serves disciplines ranging from ecology to economics, while Rényi entropy has become essential in reliability engineering, life testing, and information theory.
Life-tests typically involve constraints on overall testing duration, necessitating censoring schemes to reduce time and cost. In such schemes, the experiment terminates before all units fail. This study employs progressive Type-II censoring, where the stopping rule depends on a pre-specified number of observed failures. In this scheme, n units are placed on test at time 0, with m failures to be observed. At the first failure X ( 1 ) , R 1 surviving units are randomly withdrawn; at the second failure X ( 2 ) , R 2 units are withdrawn, and so on. At the m t h failure, all remaining units R m = n i = 1 m 1 R i m are removed. The observed data consist of the pairs ( X ( i ) , R i ) with pre-specified removal values. Although more complex to design and analyze than traditional censoring, this scheme offers greater cost and experimental efficiency while maintaining the desired estimation accuracy.
The inverse Gaussian distribution (IG) has attracted considerable attention due to its usefulness in statistical analysis, see Figure 1. The origin of the IG dates back to Tweedie [7] and Wald [8] through studies on Brownian motion. Later, Tweedie [9,10,11] provided a more detailed development of this distribution; with scale parameter λ and location parameter (mean) μ , the following probability density function (PDF), cumulative density function (CDF), survival function R ( x ) and hazard rate function r ( x ) are defined, respectively, as follows:
f ( x ; μ , λ ) = λ 2 π x 3 e λ ( x μ ) 2 2 μ 2 x , μ , λ > 0
F ( x ; μ , λ ) = Φ λ x x μ 1 + e 2 λ μ Φ λ x x μ + 1
R ( x ) = 1 Φ λ x x μ 1 + e 2 λ μ Φ λ x x μ + 1
r ( x ) = λ 2 π x 3 e λ ( x μ ) 2 2 μ 2 x 1 Φ λ x x μ 1 + e 2 λ μ Φ λ x x μ + 1
where Φ denotes the CDF of a standard normal distribution and can be expressed using the error function Φ ( x ) = 1 2 1 + e r f x 2 . The HRF of IG decreases in x for λ , μ > 0 .
Due to the extensive use of the IG distribution, numerous authors have examined the estimation of its parameters and related properties under various censoring schemes. Nath [12] introduced generalized censored samples and obtained maximum likelihood estimates. Whitmore [13] employed the EM algorithm for progressively right-censored data. Anaya and O’Reilly [14] assessed goodness-of-fit under censoring. Ismail and Auda [15] proposed kernel estimation under Type-II censoring. Jia et al. [16] developed progressive hybrid Type-I censoring and derived MLE and Bayesian estimators. Rostamian and Nematollahi [17] assessed stress-strength reliability under progressive Type-II censoring. Jayalath and Chhikara [18] performed Bayesian and fiducial survival analysis using Gibbs sampling for various censoring types. Roy et al. [19] considered progressive Type-I interval censoring. Bera and Jana [20] examined stress-strength reliability with bootstrap and Bayes estimators. Xavier et al. [21] formulated goodness-of-fit tests based on a fixed-point characterization for complete and right-censored data.
This section reviews entropy estimation under various censoring schemes. Kang et al. [22] and Cho et al. [23,24] estimated Shannon entropy for double exponential, Rayleigh, and Weibull distributions under different censoring types. Anis and De [25] studied Shannon and Rényi entropies for the Unit-Gompertz distribution. Hashem and Alyami [26] analyzed entropies for the exponential doubly Poisson distribution. Okasha and Nassar [27] and Shrahili et al. [28] estimated entropy for inverse Weibull and log-logistic distributions under progressive Type-II censoring. Maiti et al. [29] compared MLE and Bayesian estimates for generalized exponential distributions. Patra et al. [30] proposed improved entropy estimators for exponential distributions. Ahmed et al. [31] studied gamma distributions under progressive Type-I censoring, while Gong et al. [32] considered generalized inverse exponential distributions under Stepwise Type-II truncated samples.
Despite the numerous studies addressing entropy estimation for lifetime distributions under various censoring schemes, these efforts have primarily focused on distributions such as the Weibull, exponential, generalized exponential, and log-logistic models. To the best of our knowledge, no previous study has investigated the estimation of Shannon and Rényi entropies for the IG distribution under progressive Type-II censoring, despite the importance of this distribution in reliability analysis and degradation modeling, and the flexibility of this censoring scheme in life-testing experiments. This gap is particularly consequential given that the IG distribution poses well-documented estimation challenges. Folks and Chhikara [33] reported instability of maximum likelihood estimators for its parameters, particularly under small sample sizes or extreme parameter values, while Giner and Smyth [34] highlighted numerical difficulties in handling the IG likelihood within the statmod package in R, including issues related to flatness and numerical optimization. How these challenges may propagate or become amplified when estimating functions of parameters such as entropy measures has remained unexplored, and constitutes the starting point of the present study. To address this gap, this paper presents a comprehensive computational Bayesian framework that combines maximum likelihood estimation with Bayesian inference under multiple loss functions, employing Lindley’s approximation, importance sampling, and Markov chain Monte Carlo techniques. The primary contributions of this work can be summarized as follows:
  • Deriving closed-form Shannon and Rényi entropy expressions for the IG distribution, and developing a unified Bayesian computational framework under progressive Type-II censoring.
  • Documenting the breakdown of maximum likelihood estimation for Shannon entropy under certain parameter configurations, and providing practical, validated guidance for reliable entropy inference using Monte Carlo Markov Chain and importance sampling.
In this paper, we address the estimation of Shannon and Rényi entropies for the IG distribution under progressive Type-II censoring. We begin by deriving the Shannon entropy and the Rényi entropy of the IG distribution by substituting in Equation (1) and Equation (2), respectively, using Equation (3) as the basis for substitution (detailed in Supplementary Materials Section S1):
H = 0 f ( x ) ln λ 2 π x 3 e λ ( x μ ) 2 2 μ 2 x d x = 1 2 ln λ 2 π + 3 2 E ln ( x ) + λ 2 μ 2 E ( x μ ) 2 x = 1 2 ln 2 π e λ + 3 2 ln μ 3 2 e 2 λ μ E 1 2 λ μ
and
H α = 1 1 α ln 0 λ 2 π x 3 e λ ( x μ ) 2 2 μ 2 x α d x = 1 1 α ln 0 λ 2 π α 2 x 3 α 2 e α λ ( x μ ) 2 2 μ 2 x d x = 1 1 α α 2 ln λ 2 π + α λ μ + ln 2 + 1 3 α 2 ln μ + ln K 1 3 α 2 α λ μ
where K ( . ) is the Bessel function and E 1 ( . ) is the exponential integral.
The remainder of this paper is organized as follows. In Section 2, we obtain maximum likelihood estimators for the Shannon and Rényi entropies and the confidence interval (CIs). Section 3 presents Bayesian estimates using different loss functions: the squared error loss function (SE), the general entropy loss function (GE), and the LINEX loss function. Due to the absence of clear formulas for the Bayes estimates, we employ Lindley’s approximation method in Section 4, while in Section 5, we use importance sampling (IS) to derive credible intervals. In Section 6, the Monte Carlo Markov Chain (MCMC) is used to determine credible intervals and the highest posterior density (HPD). The simulation study is analyzed in Section 7 with the attachment of its resulting tables and figures and commentary on the tables. As for Section 8, it presents the real data and attaches the resulting tables and figures along with their interpretation. Finally, we review the recommendations based on the simulation results and real data, as they are linked to a previous study, to illustrate the advantages.

2. Maximum Likelihood Estimation

Let X = ( X 1 , X 2 , , X m ) denote the progressive Type-II censored sample, with X i = X i : m : n representing the i t h failure time. Although the MLEs for the parameters of the IG distribution under progressive Type-II censoring have been derived previously, we briefly restate the derivation here because the subsequent entropy estimators depend directly on these MLEs through the invariance property. The likelihood function is given by:
L ( μ , λ | x ) = C i = 1 m f ( x ( i ) ; μ , λ ) 1 F ( x ( i ) ; μ , λ ) R i
By using Equations (3) and (4) and substituting them into Equation (9), we obtain:
L ( μ , λ | x ) λ 2 π m 2 i = 1 m 1 x i 3 e i = 1 m B i i = 1 m A i R i
where A i = 1 Φ λ x i x i μ 1 + e 2 λ μ Φ λ x i x i μ + 1 , B i = λ ( x i μ ) 2 2 μ 2 x i .
The log-likelihood function is expressed as:
( μ , λ | x ) m 2 ln λ 2 π + i = 1 m ln 1 x i i = 1 m B i + i = 1 m R i ln A i
We obtain the MLEs of the unknown parameters μ and λ by maximizing the log-likelihood function given in Equation (11) with respect to μ and λ , respectively. Differentiating Equation (11) with respect to μ , and then equating the result to 0, we obtain the normal equation for μ , which is given by:
i = 1 m B ^ i , μ + i = 1 m R i A ^ i 1 Φ λ ^ x i μ ^ 2 ( e 2 λ ^ μ ^ 1 ) + e 2 λ ^ μ ^ ( 2 λ ^ x i μ ^ 3 2 λ ^ μ ^ 2 x i ) = 0
Similarly, the normal equation of λ is as follows:
m 2 λ ^ i = 1 m B ^ i , λ + i = 1 m R i A ^ i Φ 1 2 λ ^ x i μ ^ 1 x i + e 2 λ ^ μ ^ 2 λ ^ μ ^ + 1 2 λ ^ 1 x i x i μ ^ = 0
where B i , λ = B i λ , B i , μ = B i μ . We simultaneously solve Equations (12) and (13) for μ and λ to determine their maximum likelihood estimates (MLEs). Nevertheless, explicit solutions for these equations are nonexistent, indicating that the maximum likelihood estimators cannot be articulated in a closed form. We employ numerical methods to ascertain solutions to Equations (12) and (13). Let ( μ ^ , λ ^ ) denote the joint maximum likelihood estimators (MLEs) of μ and λ . The Shannon and Rényi entropy measures are nonlinear functions of ( μ , λ ) . By substituting the joint MLEs ( μ ^ , λ ^ ) into these functions, we obtain the corresponding profile maximum likelihood estimators (Pr. MLEs) of the entropies as follows:
H ^ = 1 2 ln 2 π e λ ^ + 3 2 ln μ ^ 3 2 e 2 λ ^ / μ ^ E 1 2 λ ^ μ ^
and
H ^ α = 1 1 α α 2 ln λ ^ 2 π + 1 3 α 2 ln μ ^ + ln 2 + α λ ^ μ ^ + ln K 1 3 α 2 α λ ^ μ ^
It is important to emphasize that these are profile MLEs (plug-in estimates) of the entropy measures, not true MLEs. This distinction is critical because profile MLEs of nonlinear functions may not inherit the optimal properties of MLEs and can exhibit bias and instability, particularly under censoring and for complex entropy measures such as Shannon entropy.
To compute the CIs for unknown parameters μ and λ , we utilize the asymptotic characteristics of the MLEs. In numerous real scenarios, precisely estimating these predicted values can prove challenging. We typically adopt a more straightforward method: we eliminate the expectation and compute the second derivatives directly, subsequently substituting the estimated values of μ and λ . This provides us with an estimated representation of the Fisher information matrix:
I ( μ ^ , λ ^ ) 2 ln L μ 2 2 ln L μ λ 2 ln L λ μ 2 ln L λ 2 μ = μ ^ , λ = λ ^
which we used to get an approximate variance-covariance matrix for the estimators. That helps us to understand how much uncertainty there is in each estimate and how the two parameters might be statistically related. The variance-covariance matrix of IG estimators is calculated as follows:
I ( μ ^ , λ ^ ) = ^ μ μ ^ μ λ ^ λ μ ^ λ λ
The inverse of the matrix:
I 1 = 1 ^ μ μ ^ λ λ ( ^ μ λ ) 2 ^ λ λ ^ λ μ ^ μ λ ^ μ μ = σ ^ μ μ σ ^ μ λ σ ^ λ μ σ ^ λ λ
The second derivatives with respect to the parameters μ and λ are computed (Section S2),
μ μ = i = 1 m B i , μ μ + i = 1 m R i A i , μ μ A i A i , μ 2 A i 2 , λ λ = C λ λ i = 1 m B i , λ λ + i = 1 m R i A i , λ λ A i A i , λ 2 A i 2 , μ λ = i = 1 m B i , μ λ + i = 1 m R i A i , μ λ A i A i , μ A i , λ A i 2 , λ μ = C λ μ i = 1 m B i , λ μ + i = 1 m R i A i , λ μ A i A i , λ A i , μ A i 2 .
where C = m 2 ln λ 2 π . Subsequently, we substitute in Matrix (16). Furthermore, we utilize the delta approach to obtain estimated confidence intervals for the Shannon and Rényi entropy metrics. For a comprehensive elucidation of this strategy, we direct the reader to Greene [35]. Assume that
τ H T = H ( μ , λ ) μ , H ( μ , λ ) λ and τ H α T = H α ( μ , λ ) μ , H α ( μ , λ ) λ .
Moreover, the approximate estimated variances in H ^ and H ^ α were given by
Var ^ ( H ^ ) = τ H T I ^ ( μ ^ , λ ^ ) 1 τ H and Var ^ ( H ^ α ) = τ H α T I ^ ( μ ^ , λ ^ ) 1 τ H α .
Therefore, ( H ^ H ) / Var ^ ( H ^ ) and ( H ^ α H α ) / Var ^ ( H ^ α ) asymptotically followed the standard normal distribution. Thus, the asymptotic CIs (Wald interval) 100 ( 1 p ˜ ) % for the Shannon and Rényi entropies were given, respectively, by
H ^ ± Z p ˜ / 2 Var ^ ( H ^ ) and H ^ α ± Z p ˜ / 2 Var ^ ( H ^ α ) .
where Z p ˜ / 2 denoted the upper ( p ˜ / 2 ) t h percentile of the standard normal distribution. The final result is obtained numerically because the equations involved are complex to solve analytically. It is important to note that the Wald-CIs constructed via the delta method are asymptotic in nature, and their finite sample performance depends on the shape of the likelihood function, the nonlinearity of the entropy formula, and the stability of the variance-covariance matrix. It is well documented that Wald intervals can perform poorly when the likelihood surface is irregular or the sample size is small, as exemplified by the binomial success probability [36].
Although the MLEs are straightforward to compute via the invariance principle, their finite-sample performance under progressive censoring (particularly for the entropy functions) remains to be evaluated. This motivates the development of Bayesian alternatives in the following section.

3. Bayesian Estimation

After deriving the MLEs of the uncertainty measures in Equations (7) and (8) in the preceding section, we now proceed to Bayesian inference for the Shannon and Rényi entropy measures. This Bayesian approach commences with the definition of a prior distribution and a likelihood function, which collectively yield the posterior distribution of the quantities of interest. Choosing an appropriate prior distribution for unknown parameters is a fundamental challenge in Bayesian inference. In the absence of strong prior knowledge, a gamma distribution was adopted as the prior for both μ and λ , consistent with prior specifications previously employed in the context of the IG distribution [37]. The specific numerical values of the prior hyperparameters, along with the corresponding sensitivity analysis, are detailed in the simulation section, where practical performance is assessed under multiple prior scenarios.
A random variable X is considered to adhere to a gamma distribution, assuming independent gamma priors for μ and λ , respectively, in this context:
g 1 ( μ ; a , b ) = b a μ a 1 e μ b Γ ( a ) , g 2 ( λ ; c , d ) = d c λ c 1 e λ d Γ ( c ) . λ > 0 , a , b , c , d > 0
We give the joint prior distribution for μ and λ as:
g ( μ , λ ) = b a d c Γ ( a ) Γ ( c ) ( μ a 1 ) ( λ c 1 ) e ( μ b + λ d ) μ , λ > 0 , a , b , c , d > 0
We take the log-joint prior distribution for μ and λ as:
ρ ( μ , λ ) = a ln ( b ) + c ln ( d ) + ( a 1 ) ln ( μ ) + ( c 1 ) ln ( λ ) μ b λ d ln Γ ( a ) ln Γ ( c )
where ρ ( μ , λ ) = ln g ( μ , λ ) . After making calculations using Equations (10) and (17), the joint posterior distribution of μ and λ can be established as follows:
Π ( μ , λ | x ̲ ) = 1 z ˜ ( μ a 1 ) ( λ m 2 + c 1 ) ( e ( μ b + λ d ) ) i = 1 m 1 x i 3 e B i A i R i
where z ˜ is a normalizing constant defined by,
z ˜ = 0 0 ( μ a 1 ) ( λ m 2 + c 1 ) ( e ( μ b + λ d ) ) i = 1 m 1 x i 3 e B i A i R i d μ d λ
The posterior density function of μ , λ is determined by Bayes’ method:
Π ( μ , λ | x ̲ ) = L ( μ , λ ) g ( μ , λ ) ( μ , λ ) L ( μ , λ ) g ( μ , λ ) d ( μ , λ )
This paper analyzes three forms of loss function: the SE loss function, the GE loss function, and the LINEX loss function. Let H ^ denote an estimator for the entropy functions H. Subsequently, we develop the loss functions, as illustrated in Table 1.
The Bayes estimator under loss function can be expressed as:
H ^ B = T 1 E ð ( H ) | x ̲ = T 1 ( μ , λ ) ð ( H ) Π ( μ , λ | x ̲ ) d ( μ , λ )
where ð ( H ) is the transformation applied to the entropy before taking the posterior expectation, and T 1 is the inverse operation that recovers H ^ B from the expected transformed value. For SE loss function, ð ( H ) = H and T 1 is the identity. For GE loss function, ð ( H ) = H q and T 1 ( x ) = x 1 / q . For LINEX loss function, ð ( H ) = e p H and T 1 ( x ) = 1 p ln ( x ) .
Using Equation (21) and Equation (19), the Bayes estimates of the Shannon entropy H with respect to the GE and LINEX loss functions are obtained, respectively, as follows:
H = 1 z ˜ 0 0 ( H ) q ( μ a 1 ) ( λ m 2 + c 1 ) ( e ( μ b + λ d ) ) i = 1 m 1 x i 3 A i R i d μ d λ 1 q
H = 1 p ln 1 z ˜ 0 0 ( μ a 1 ) ( λ m 2 + c 1 ) ( e ( μ b + λ d + p H ) ) i = 1 m 1 x i 3 A i R i d μ d λ
By replacing q = 1 into Equation (22), we obtain the Bayes estimate of Shannon entropy H concerning the SE loss function. Likewise, we employ Equation (21) and Equation (19) to derive the Bayes estimates of the Rényi entropy H α based on the aforementioned loss functions.
Computing closed-form solutions for the ratio of the two integrals presented in Equations (22) and (23) is intricate. To tackle this issue, we utilize Lindley’s approximation technique.

4. Lindley’s Approximation

In this section, we derive Lindley’s approximation for Bayesian estimates, which depend on two parameters, μ and λ . The main equation was proposed by Lindley [38]. It is expressed as follows:
I ( x ) = E [ u ( μ , λ | x ̲ ) ] = ( μ , λ ) u ( μ , λ ) e ( μ , λ ) + ρ ( μ , λ ) d ( μ , λ ) ( μ , λ ) e ( μ , λ ) + ρ ( μ , λ ) d ( μ , λ )
where
u ( μ , λ ) = The objective function of μ and λ , ( μ , λ ) = The log of likelihood function , ρ ( μ , λ ) = The log of joint prior of μ and λ .
can be evaluated as
I ( x ) = u ( μ ^ , λ ^ ) + 1 2 ( u ^ λ λ + 2 u ^ λ ρ ^ λ ) σ ^ λ λ + ( u ^ μ λ + 2 u ^ μ ρ ^ λ ) σ ^ μ λ + ( u ^ λ μ + 2 u ^ λ ρ ^ μ ) σ ^ λ μ + ( u ^ μ μ + 2 u ^ μ ρ ^ μ ) σ ^ μ μ + 1 2 [ ( u ^ λ σ ^ λ λ + u ^ μ σ ^ λ μ ) ( ^ λ λ λ σ ^ λ λ + ^ λ μ λ σ ^ λ μ + ^ μ λ λ σ ^ μ λ + ^ μ μ λ σ ^ μ μ ) + ( u ^ λ σ ^ μ λ + u ^ μ σ ^ μ μ ) ( ^ μ λ λ σ ^ λ λ + ^ λ μ μ σ ^ λ μ + ^ μ λ μ σ ^ μ λ + ^ μ μ μ σ ^ μ μ ) ]
where
μ ^ = MLE of μ , λ ^ = MLE of λ , u ^ λ = u ( μ ^ , λ ^ ) λ ^ , u ^ μ = u ( μ ^ , λ ^ ) μ ^ , u ^ λ λ = 2 u ( μ ^ , λ ^ ) λ ^ 2 , u ^ μ μ = 2 u ( μ ^ , λ ^ ) μ ^ 2 , u ^ λ μ = 2 u ( μ ^ , λ ^ ) λ ^ μ ^ , u ^ μ λ = 2 u ( μ ^ , λ ^ ) μ ^ λ ^ , ^ λ λ μ = 3 ( μ ^ , λ ^ ) λ ^ 2 μ ^ , ^ λ λ λ = 3 ( μ ^ , λ ^ ) λ ^ 3 , ^ λ μ λ = 3 ( μ ^ , λ ^ ) λ ^ μ ^ λ ^ , ^ μ μ λ = 3 ( μ ^ , λ ^ ) μ ^ 2 λ ^ , ^ μ λ λ = 3 ( μ ^ , λ ^ ) μ ^ λ ^ 2 , ^ λ μ μ = 3 ( μ ^ , λ ^ ) λ ^ μ ^ 2 , ^ μ μ μ = 3 ( μ ^ , λ ^ ) μ ^ 3 , ^ μ λ μ = 3 ( μ ^ , λ ^ ) μ ^ λ ^ μ ^ , ρ ^ λ = ρ ( μ ^ , λ ^ ) λ ^ , ρ ^ μ = ρ ( μ ^ , λ ^ ) μ ^ .
That approximation method has been employed by numerous researchers, including Nassar et al. [39], Kundu and Pradhan [40], Xu et al. [41], Kim et al. [42], Sultan and Ahmad [43], Sana and Faizan [44], and Ren and Hu [45].
⇒ Firstly; 
To apply Lindley’s approximation, the function u ( μ , λ ) appearing in the posterior expectation must be specified for each loss function. From the general form given in Equation (21), we obtain the following (detailed derivations and the required partial derivatives are provided in Section S2):
  • Bayesian estimators of the Shannon entropy under the SE, GE, and LINEX loss function, respectively, as follows:
    u SE ( μ , λ ) = H , u GE ( μ , λ ) = H q and u LINEX ( μ , λ ) = e H p .
  • Bayesian estimators of the Rényi entropy under the SE, GE, and LINEX loss function, respectively, as follows:
    u SE ( μ , λ ) = H α , u GE ( μ , λ ) = H α q and u LINEX ( μ , λ ) = e H α p .
⇒ Secondly; 
we obtain the derivatives of the ML equations (in Section S2):
^ μ μ μ = i = 1 m B i , μ μ μ + i = 1 m R i A i , μ μ μ A i 2 3 A i , μ A i , μ μ A i + 2 A i , μ 3 A i 3 ^ μ μ λ = i = 1 m B i , μ μ λ + i = 1 m R i A i , μ μ λ A i 2 A i A i , μ μ A i , λ 2 A i A i , μ A i , μ λ + 2 A i , μ 2 A i , λ A i 3 ^ μ λ μ = i = 1 m B i , μ λ μ + i = 1 m R i A i , μ λ μ A i 2 2 A i A i , μ A i , μ λ A i A i , μ μ A i , λ + 2 A i , μ 2 A i , λ A i 3 ^ μ λ λ = i = 1 m R i A i , μ λ λ A i 2 2 A i A i , μ λ A i , λ A i A i , μ A i , λ λ + 2 A i , μ A i , λ 2 A i 3 ^ λ λ λ = C λ λ λ + i = 1 m R i A i , λ λ λ A i 2 3 A i , λ A i , λ λ A i + 2 A i , λ 3 A i 3 ^ λ λ μ = i = 1 m R i A i , λ λ μ A i 2 A i , λ λ A i , μ A i 2 A i , λ μ A i , λ A i + 2 A i , λ 2 A i , μ A i 3 ^ λ μ λ = i = 1 m R i A i , λ λ μ A i 2 A i , λ λ A i , μ A i 2 A i , λ μ A i , λ A i + 2 A i , λ 2 A i , μ A i 3 ^ λ μ μ = i = 1 m B i , λ μ μ + i = 1 m R i A i , λ μ μ A i 2 A i , λ A i , μ μ A i 2 A i , λ μ A i , μ A i + 2 A i , λ A i , μ 2 A i 3
⇒ Thirdly; 
we obtain the derivatives of the log-joint prior distribution from Equation (18):
ρ ^ μ = a 1 μ b , ρ ^ λ = c 1 λ d .
⇒ Fourthly; 
we substitute it into Equation (25) to obtain the result. We change the objective function ( u ( μ ^ , λ ^ ) ) and its derivatives ( u ^ λ , u ^ μ , u ^ λ λ , u ^ μ μ , u ^ λ μ and u ^ μ λ ) each time due to the different loss functions every time, while keeping the other derivatives related to the maximum likelihood function ( ^ μ μ μ , ^ μ μ λ , ^ μ λ μ , ^ μ λ λ , ^ λ λ λ , ^ λ λ μ , ^ λ μ λ and ^ λ μ μ ), Fisher matrix elements ( σ ^ λ λ , σ ^ μ λ , σ ^ λ μ , and σ ^ μ μ ) and also the derivatives of joint prior distribution ( ρ ^ λ and ρ ^ μ ) fixed.
Thus, the Bayes estimate of Shannon entropy H with respect to the GE loss function is obtained as;
[ E ( [ H ] q | x ̲ ) ] 1 q = [ H q + 1 2 [ ( u ^ λ λ + 2 u ^ λ ρ ^ λ ) σ ^ λ λ + ( u ^ μ λ + 2 u ^ μ ρ ^ λ ) σ ^ μ λ + ( u ^ λ μ + 2 u ^ λ ρ ^ μ ) σ ^ λ μ + ( u ^ μ μ + 2 u ^ μ ρ ^ μ ) σ ^ μ μ ] + 1 2 [ ( u ^ λ σ ^ λ λ + u ^ μ σ ^ λ μ ) ( ^ λ λ λ σ ^ λ λ + ^ λ μ λ σ ^ λ μ + ^ μ λ λ σ ^ μ λ + ^ μ μ λ σ ^ μ μ ) + ( u ^ λ σ ^ μ λ + u ^ μ σ ^ μ μ ) ( ^ μ λ λ σ ^ λ λ + ^ λ μ μ σ ^ λ μ + ^ μ λ μ σ ^ μ λ + ^ μ μ μ σ ^ μ μ ) ] ] 1 q
By substituting q = 1 in Equation (26), we can obtain the Bayes estimate of H with respect to the SE loss function. Using a similar procedure, we can obtain the Bayes estimates of the Rényi entropy H α with respect to the above loss functions.
Therefore, the Bayes estimate of Shannon entropy H with respect to the LINEX loss function is obtained as;
[ E ( e [ H ] p | x ̲ ) ] = 1 p ln [ e H p + 1 2 [ ( u ^ λ λ + 2 u ^ λ ρ ^ λ ) σ ^ λ λ + ( u ^ μ λ + 2 u ^ μ ρ ^ λ ) σ ^ μ λ + ( u ^ λ μ + 2 u ^ λ ρ ^ μ ) σ ^ λ μ + ( u ^ μ μ + 2 u ^ μ ρ ^ μ ) σ ^ μ μ ] + 1 2 [ ( u ^ λ σ ^ λ λ + u ^ μ σ ^ λ μ ) ( ^ λ λ λ σ ^ λ λ + ^ λ μ λ σ ^ λ μ + ^ μ λ λ σ ^ μ λ + ^ μ μ λ σ ^ μ μ ) + ( u ^ λ σ ^ μ λ + u ^ μ σ ^ μ μ ) ( ^ μ λ λ σ ^ λ λ + ^ λ μ μ σ ^ λ μ + ^ μ λ μ σ ^ μ λ + ^ μ μ μ σ ^ μ μ ) ] ]
Using a similar procedure, we can obtain the Bayes estimate of the Rényi entropy H α with respect to the above loss function by using Equation (27).
Although Lindley’s approximation provides a computationally efficient way to obtain Bayesian estimates, it does not easily produce credible intervals for entropy measures. So, we now introduce the importance sampling procedure, which makes it easy to build HPD in addition to point estimates.

5. Importance Sampling

This paper examines the importance sampling approach for calculating Bayes estimates of unknown parametric functions as defined in Equations (7) and (8). The exception in IG is that we need to set a different proposal for each parameter μ and λ , while in most classical distributions, a conjugate distribution is enough to cover everything. First, we define the proposal distribution ( p ) from which we generate the parameters related to the IG μ and λ . Note that it is not necessary for the proposal distribution to be the same as the prior distribution. We need to determine the distribution of the proposal using the joint posterior distribution of μ and λ as follows:
Π ( μ , λ | x ̲ ) μ a 1 e μ b λ m 2 + c 1 e λ d e i = 1 m B i i = 1 m 1 x i 3 A i R i
We focus on the parts that contain λ :
Π ( μ , λ | x ̲ ) λ m 2 + c 1 exp ( λ d + 1 2 i = 1 m ( x i μ ) 2 μ 2 x i λ > 0
This represents the form of the gamma distribution, so the most suitable proposal distribution ( p λ ) for λ is the gamma distribution (G).
We focus on the parts that contain μ :
Π ( μ , λ | x ̲ ) μ a 1 exp μ b λ 2 i = 1 m ( x i μ ) 2 2 μ 2 x i μ > 0
It seems that μ is more complicated because it contains this specific part ( x μ ) 2 / 2 μ 2 x , since it does not resemble the gamma distribution. Instead, the log-posterior density can be locally approximated in the neighborhood of the maximum likelihood estimator μ ^ using a second-order Taylor expansion. Consequently, the posterior distribution is approximately normal around μ ^ , which motivates the use of a normal proposal distribution. To make it clear, we follow these steps:
L ( μ ) μ a 1 exp μ b λ 2 i = 1 m ( x i μ ) 2 2 μ 2 x i
We take the logarithm:
( μ ) = ln L ( μ ) = ( a 1 ) ln ( μ ) μ b λ 2 + 1 2 μ 2 i = 1 m x i m μ + 1 2 i = 1 m 1 x i
We find the first derivative and equating it to z e r o :
d ( μ ) d μ = a 1 μ b i = 1 m x i μ 3 + m μ 2 a 1 μ b i = 1 m x i μ 3 + m μ 2 = 0 m μ i = 1 m x i = μ 2 ( b μ ( a 1 ) )
We find the second derivative to determine the maximum and minimum values:
d 2 ( μ ) d μ 2 = a 1 μ 2 + m μ 3 3 m μ i = 1 m x i μ 4
from     and  
d 2 ( μ ) d μ 2 = 2 ( a 1 ) μ 3 b μ 2 + m μ 3
If d 2 ( μ ) d μ 2 μ = μ ^ > 0 means that point is a minimum, then d 2 ( μ ) d μ 2 μ = μ ^ < 0 also means that point is a maximum (saddle point).
Looking at the previous expression 2 ( a 1 ) μ 3 b μ 2 + m / μ 3 , what determines whether it is positive or negative is the numerator 2 ( a 1 ) μ 3 b μ 2 + m , because the denominator μ 3 is always positive anyway. So, we plug in numbers into the numerator and look for the values that make it negative so the curve goes down, which means that point is a maximum. The larger the value of the quantity 3 b μ 2 , the more negative the second derivative.
Because the functional form of the proposal distribution ( p μ ) for μ is unknown, we use Laplace approximation to determine an appropriate Gaussian form, where the logarithm of the likelihood function is expanded around the maximum likelihood estimate of μ using a second-order Taylor series, which leads to a normal approximation of the proposal distribution given by μ N μ ^ , [ I ( μ ^ ) ] 1 , where I ( μ ^ ) denotes the observed Fisher information. This approximation provides a convenient and efficient proposal distribution for the importance sampling procedure.
However, since μ > 0 , it is more appropriate to propose a truncated normal distribution because the standard normal distribution extends from [ , ] , so if μ N ( μ 0 , σ 2 ) , it could take a positive or negative value, which would be illogical in this context. Therefore, μ Trunc . Normal ( μ 0 , σ 2 , A , B ) where [ A , B ] represents the truncation limits (that is, values outside the interval are not possible). In this approach, it is necessary to reformulate the joint posterior distribution of μ and λ as presented in Equation (19). It is provided by
Π ( μ , λ | x ̲ ) G λ m 2 + c , d + 1 2 i = 1 m ( x i μ ) 2 μ 2 x i N μ | λ μ ^ , [ I ( μ ^ ) ] 1 [ A , B ] w ( μ , λ )
For each pair generated ( μ i , λ i ) , the importance weight is computed as
w i = L ( x μ i , λ i ) g ( μ i , λ i ) p μ ( μ i ) p λ ( λ i )
where L ( x μ i , λ i ) g ( μ i , λ i ) = Posterior ( μ i , λ i ) (before normalization). The normalized weights are then obtained as
w ˜ i = w i j = 1 N w j
Later, we will use these weights to calculate Bayesian estimates of any parametric function h ( μ , λ ) as
I ^ weight = i = 1 N w ˜ i h ( μ i , λ i )
The following procedure can be used to obtain Bayes estimates of a function, denoted as g ( μ , λ ) , with respect to the GE and LINEX loss functions:
Step 1.
Generate λ from a gamma distribution with shape parameter m 2 + c and scale parameter d + 1 2 i = 1 m ( x i μ ) 2 μ 2 x i , denoted as G λ m 2 + c , d + 1 2 i = 1 m ( x i μ ) 2 μ 2 x i .
Step 2.
For the given λ obtained in Step 1., approximate the MLE round for μ , the form approaching N μ | λ μ ^ , [ I ( μ ^ ) ] 1 [ A , B ] .
Step 3.
Execute Step 1. and Step 2. for N iterations to derive ( μ 1 , λ 1 ) , ( μ 2 , λ 2 ) , , ( μ N , λ N ) .
Step 4.
The Bayes estimates of the parametric function h ( μ , λ ) under the GE and LINEX loss functions are derived, respectively, as follows:
I ^ GE = i = 1 N h ( μ i , λ i ) q w ( μ i , λ i ) 1 q i = 1 N w ( μ i , λ i ) 1 q
and
I ^ LINEX = 1 p ln i = 1 N e p h ( μ i , λ i ) w ( μ i , λ i ) i = 1 N w ( μ i , λ i ) 1
The Bayes estimate of h ( μ , λ ) concerning the SE loss function I ^ SE can be obtained from Equation (28) for q = 1 . Furthermore, within the framework of the GE loss function, Bayesian estimates of the Shannon and Rényi entropy functions can be derived from Equation (28). Similarly, for the LINEX loss function, Bayesian estimates of the Shannon and Rényi entropy functions can be obtained from Equation (29), where h ( μ , λ ) = H corresponds to the Shannon entropy and h ( μ , λ ) = H α pertains to the Rényi entropy.
Step 5.
Propose intervals for the estimation of the Shannon entropy function and the Rényi entropy function. This method was used by Chen and Shao [46] for this purpose. Define
η i = η ( μ ( i ) , λ ( i ) )
and
w i = Posterior ( μ ( i ) , λ ( i ) ) p ( μ ( i ) , λ ( i ) ) w ˜ i = w ( μ ( i ) , λ ( i ) ) i = 1 N w ( μ ( i ) , λ ( i ) )
where λ ( i ) and μ ( i ) for i = 1 , 2 , , N are posterior samples generated from previous steps for λ and μ , respectively. Each sample represents a part of the posterior distribution, and the size of that part (its proportion) is given by the normalized weight, w ˜ i .
Step 6.
Estimate the γ t h quantile of η i = η ( μ ( i ) , λ ( i ) ) :
η ^ ( γ ) = η ( 1 ) , γ = 0 η ( i ) , j = 1 i 1 w ˜ ( j ) < γ j = 1 i w ˜ ( j )
Let η ( i ) be the ordered values of η i , and let w ˜ ( j ) be the corresponding normalized weights for these ordered values.
Step 7.
Establish the typical credible interval with a 100 ( 1 p ˜ ) % confidence level: Let p be the target coverage probability (the confidence level), where p = 1 p ˜ . For a 95% credible interval, p = 0.95 and p ˜ = 0.05 . The bottom bound of the credible interval is the p ˜ / 2 quantile of the posterior distribution of η . The top limit of the credible interval is the 1 p ˜ / 2 quantile of the posterior distribution of η . The standard credible interval, sometimes referred to as the equal-tail interval, is computed utilizing the quantile estimations from Equation (30).
η ^ ( p ˜ / 2 ) , η ^ ( 1 p ˜ / 2 )
where, η ^ ( p ˜ / 2 ) is the estimated quantile below which 100 ( p ˜ / 2 ) % of the posterior distribution of the entropy lies. η ^ ( 1 p ˜ / 2 ) is the estimated quantile below which 100 ( 1 p ˜ / 2 ) % of the posterior distribution lies.
Although importance sampling provides accurate estimates when the proposal distribution is well-chosen, it does not automatically explore the full posterior. We therefore also employ MCMC, which constructs a Markov chain targeting the exact posterior and serves as the gold-standard simulation-based method in this study.

6. Monte Carlo Markov Chain

MCMC algorithms offer a versatile approach to create approximate samples from any desired distribution. They are extensively utilized in contemporary statistics, particularly in Bayesian analysis. In Bayesian inference, the posterior distribution is the primary emphasis; nevertheless, it frequently cannot be expressed in a closed form, as is the case here. Two prevalent MCMC methods utilized in Bayesian estimation include the Metropolis–Hastings algorithm, presented by Hastings [47], and the Gibbs sampler, proposed by Geman and Geman [48]. Both methodologies utilize marginal posterior distributions throughout the sampling procedure. In this approach, it is necessary to reformulate the joint posterior distribution of μ and λ as presented in Equation (19). It is provided by
Π ( μ , λ | x ̲ ) μ a 1 λ m 2 + c 1 e ( μ b + λ d ) i = 1 m 1 x i 3 e B i A i R i
When examining the posterior distribution in Equation (19), the complete conditional posterior distributions for the parameter μ and the parameter λ are derived utilizing the informative gamma priors, as delineated below.    
Π ( μ | λ , x ̲ ) μ a 1 e μ b i = 1 m exp λ ( x i μ ) 2 2 μ 2 x i A i R i
and
Π ( λ | μ , x ̲ ) λ m 2 + c 1 e λ d i = 1 m exp λ ( x i μ ) 2 2 μ 2 x i A i R i
Due to the inability to derive the marginal distributions Π ( μ | λ , x ̲ ) and Π ( λ | μ , x ̲ ) in closed form, we employ the Metropolis–Hastings algorithm, which facilitates the estimation of Bayesian point estimates and credible intervals for unknown parameters μ and λ , in addition to the entropy measures H and H α . A cyclic implementation of the Metropolis–Hastings sampling process to acquire approximate samples from the posterior distribution Π . This can be acquired using the subsequent steps:
Step 1.
Specify starting values μ ( 0 ) and λ ( 0 ) (assume MLEs of μ and λ ) and specify J = 1 .
Step 2.
Generate μ * and λ * from the normal proposal distributions N μ ( J 1 ) , σ μ 2 and N λ ( J 1 ) , σ λ 2 , respectively, where σ μ 2 and σ λ 2 are the variances in μ and λ , which can be estimated from the inverse of Fisher information matrix.
Step 3.
Determine the following acceptance probabilities:
r μ = min Π ( μ * | λ ( J 1 ) , x ̲ ) Π ( μ ( J 1 ) | λ ( J 1 ) , x ̲ ) , 1 and r λ = min Π ( λ * | μ ( J 1 ) , x ̲ ) Π ( λ ( J 1 ) | μ ( J 1 ) , x ̲ ) , 1 .
Step 4.
Generate samples u 1 and u 2 from uniform distribution U ( 0 , 1 ) .
Step 5.
If u 1 r μ (acceptance probability), accept the proposal and specify μ ( J ) = μ * , otherwise μ ( J ) = μ ( J 1 ) . Similarly, if u 2 r λ , accept the proposal and specify λ ( J ) = λ * , otherwise λ ( J ) = λ ( J 1 ) .
Step 6.
Compute the Shannon entropy and Rényi entropy indicators from Equation (7) and Equation (8), respectively:
H ( J ) = 1 2 ln 2 π e λ ( J ) + 3 2 ln μ ( J ) 3 2 e 2 λ μ ( J ) E 1 2 λ ( J ) μ ( J )
and
H α ( J ) = 1 1 α α 2 ln λ ( J ) 2 π + α λ ( J ) μ ( J ) + ln 2 + 1 3 α 2 ln μ ( J ) + ln K 1 3 α 2 α λ ( J ) μ ( J )
Step 7.
Specify J = J + 1
Step 8.
Repeat Step 1. to Step 6., N MC (the overall number of iterations) times to obtain MCMC samples. H ( 1 ) , H α ( 1 ) , H ( 2 ) , H α ( 2 ) , …, H ( N MC ) , H α ( N MC ) represent the generated samples obtained from the MH algorithm. After discarding the first M number of burn-in samples, the remaining N MC M samples are used to compute Bayesian estimates of H and H α . The Bayes estimate of η = H , H α under the SE, GE, and LINEX loss functions can now be computed as:
η ^ GE = 1 N MC M j = M + 1 N MC η j q 1 q
and
η ^ LINEX = 1 p ln 1 N MC M j = M + 1 N MC e p η j
The Bayes estimate of η ^ SE with respect to the SE loss function can be derived from Equation (36) when q = 1 .
Step 9.
Order the values η ( M + 1 ) , η ( M + 2 ) , , η ( N MC ) ascendingly as η ( 1 ) < η ( 2 ) < < η ( N MC M ) . Assuming the desired confidence level is 100 ( 1 p ˜ ) % (where p ˜ is the significance level), the standard Bayesian credible interval for the parameter η is given by
η ( N MC M ) · ( p ˜ / 2 ) , η ( N MC M ) · ( 1 p ˜ / 2 )
where . denotes the integer part (floor function). η ( k ) is the k t h element in the ordered sequence. p ˜ / 2 is the left-tail probability, and 1 p ˜ / 2 is the right-tail probability. Therefore, the HPD of η can be constructed.
With all three Bayesian computational methods now defined (Lindley’s approximation, importance sampling, and MCMC). This methodological summary is illustrated in Algorithm 1, their practical performance in estimating Shannon and Rényi entropies under progressive Type-II censoring is assessed through extensive simulations in the following section.
Algorithm 1: Bayesian estimation of H(μ, λ) and Hα(μ, λ) for the IG distribution under progressive Type-II censoring
Entropy 28 00871 i001

7. Simulation Study

A comprehensive simulation study was carried out to assess the performance of the proposed estimation methods under progressive Type-II censoring of the entropy functions given by Equations (7) and (8). The study was based on 1000 Monte Carlo replications. Let n be the total sample size and m ( m n ) be the effective sample size after censoring. Progressive Type-II censored schemes are defined by the removal vector R = ( R 1 , R 2 , , R m ) , where R i units are randomly removed at the i t h failure time. Three censoring schemes (S1, S2, and S3) were examined for different sample sizes ( n = 30 , 50 , 100 ) and effective sample sizes ( m = 20 , 40 , 60 , 75 ). The schemes are summarized in Table 2. Two parameter settings were considered to evaluate performance under different distributional shapes: Setting I with μ = 1.5 , λ = 2.0 , yielding true Shannon entropy H = 1.2481 and Rényi entropy H α = 1.6894 (at α = 0.5 ); and Setting II with μ = 2.0 , λ = 1.0 , yielding true Shannon entropy H = 1.5641 and Rényi entropy H α = 2.2997 (at α = 0.5 ). For the Bayesian procedures, the informative priors μ Gamma ( 3 , 2 ) and λ Gamma ( 10 , 5 ) were adopted, yielding prior means of 1.5 and 2, respectively, consistent with the parameter values of Setting I. Notably, these prior means do not match the parameter values of Setting II, thereby providing an implicit robustness check: the Bayesian procedures were tested under a prior specification centered away from the true values in Setting II. The results confirmed the stability of the MCMC and importance sampling methods across both settings, reflecting the dominance of the likelihood over the prior.
In each replication, samples were generated from an IG distribution with parameters ( μ = 1.5 , λ = 2 ) and ( μ = 2 , λ = 1 ) using the progressive Type-II censoring algorithm presented by Balakrishnan and Sandhu [49] as follows:
Step 1.
Generate m independent Uniform (0,1) observations w 1 * , w 2 * , , w m * .
Step 2.
For a chosen censoring scheme from Table 2, set the removal vector R i for i = 1 , 2 , , m .
Step 3.
Set E i * = i + j = m i + 1 m R j 1 for i = 1 , 2 , , m .
Step 4.
Set V i * = ( w i * ) E i * for i = 1 , 2 , , m .
Step 5.
Set U i , m , n * = U i * = 1 k = m i + 1 m V k * for i = 1 , 2 , , m . Then, U 1 * , U 2 * , , U m * is the required progressive Type-II censored sample from the Uniform(0,1) distribution.
Step 6.
Finally, the progressive Type-II censored sample from the IG distribution is obtained by setting X i = F 1 ( U i * ) , where F 1 ( · ) is the inverse CDF from Equation (4), which must be solved numerically.
For the MCMC method, we ran the parameter generation chain for 10,000 iterations, discarding the first 2000 as the burn-in period. For importance sampling, we generated 5000 parameter samples and then subjected them to the specified progressive censoring scheme. For each replication, point and interval estimates were calculated. Summary measures, such as mean estimates, bias, and coverage probabilities, are obtained by averaging over the 1000 Monte Carlo runs. The simulation results are summarized in Table 3, Table 4 and Table 5 as shown:
In Table 3, Pr. MLE exhibited poor performance, with a substantial positive bias ( 0.33 0.41 ) and a high mean relative error ( 0.30 0.34 ) , showing only limited improvement as the sample size increased. In contrast, the IS, MCMC, and Lindley methods (under the GE and LINEX loss functions) produced highly accurate estimates, with low mean relative errors ( 0.07 0.14 ) and biases close to z e r o , whereas the Lindley estimator under the SE loss function was the least efficient among these superior methods. The inferior performance of the Pr. MLE can be attributed to the Shannon entropy formula, which contains the highly sensitive term e 2 λ / μ E 1 ( 2 λ / μ ) . This term amplifies parameter estimation errors arising from the complex likelihood surface of the IG distribution. By contrast, the Bayesian methods and IS overcome this limitation by relying on the entire posterior (or sampling) distribution of the entropy rather than on a single point estimate.
In Table 4, performance of the Pr. MLE improved considerably compared with the Shannon entropy results. It yielded a smaller negative bias 0.02 to 0.09 and an acceptable mean relative error ( 0.074 0.174 ) , with clear improvement as the sample size increased. Nevertheless, the IS, MCMC, and Lindley methods remained more accurate, achieving MRE values ranging from 0.068 to 0.126 . However, the differences among all estimation methods became less pronounced, and the performance of the different loss functions within the Lindley approach was remarkably similar. This improvement can be explained by the mathematical form of Rényi entropy, which depends on the modified Bessel function K ν . This function is smoother and less prone to magnifying estimation errors, while the parameter moderates the influence of variations in the ratio λ / μ , making the entropy less sensitive to parameter estimation errors.
In Table 5, Wald-CIs consistently produced excessive coverage (with the coverage probability often reaching 1.000), but at the expense of excessively wide and impractical interval lengths. For example, the interval length reached 4.33 for Shannon entropy when n = 30 and remained approximately three to four times longer than those of the competing methods even when n = 100 . In contrast, the IS and MCMC approaches (using both equal-tail confidence intervals and HPD credible intervals) generated substantially shorter intervals ( 0.46 0.53 for Shannon entropy when n = 100 ) while maintaining coverage probabilities close to the nominal 0.95 level. The excessive width of the Wald intervals reflects the instability of the covariance matrix caused by the irregular likelihood surface. In comparison, the simulation-based methods produce more efficient interval estimates because they are constructed from the full empirical distribution of the entropy, thereby capturing the dependence structure between the model parameters. It is also noteworthy that the interval lengths for Rényi entropy were consistently shorter than those for Shannon entropy across all methods, further confirming the greater numerical stability of the Rényi entropy formulation.
The simulation results are summarized in Table 6, Table 7 and Table 8 as shown:
In Table 6, Pr. MLE exhibited a positive bias ranging from 0.74 to 0.90 and a mean relative error between 0.49 and 0.58, both of which were higher than those observed under the first set of simulation settings. The Lindley method showed some variability in performance, with relatively large mean squared error values at small sample sizes, particularly under Scenario S1, although its accuracy improved noticeably as the sample size increased. In contrast, the IS and MCMC methods produced more accurate estimates, with mean relative errors ranging from 0.067 to 0.12 and biases closer to z e r o , while maintaining stable performance across all simulation scenarios.
In Table 7, Pr. MLE performed relatively better than it did for Shannon entropy, yielding a negative bias ranging from −0.015 to −0.126 and a mean relative error between 0.077 and 0.207. The stability issues observed for the Lindley method were limited to the SE loss function at small sample sizes, whereas the performance of the different loss functions became increasingly similar as the sample size grew. The IS and MCMC methods maintained consistently stable performance, with mean relative errors ranging from 0.068 to 0.113 and no notable outlying values.
In Table 8, for Shannon entropy, the Wald-CIs were relatively wide, and their coverage probability declined as the sample size increased, reaching as low as 0.500 in some scenarios. In contrast, the Wald-CIs for Rényi entropy maintained relatively high coverage probabilities but remained wider than the corresponding intervals produced by the competing methods. The IS and MCMC interval estimators achieved coverage probabilities close to the nominal level in most cases while producing considerably shorter intervals than the Wald-CIs, although slight reductions in coverage were observed for certain sample sizes. Overall, the interval estimation results for Rényi entropy were more stable than those for Shannon entropy across all estimation methods.
The results of Figure 2, Figure 3, Figure 4 and Figure 5 show the effectiveness of the MCMC approach for a variety of parameters and entropy measures settings. This suggests that the MCMC approach is a powerful and reliable method to sample complex distributions.

8. Real Data

Across both Dataset I and Dataset II, the IG distribution consistently demonstrated superior performance compared to the following distributions based on both the AIC criterion and the Kolmogorov–Smirnov goodness-of-fit test results; Epanechnikov–Weibull (EpW) distribution; Perks (Perk) distribution; Gompertz–Lindley (GomLin) distribution; Gamma (Gam) distribution; Power Muth (PMuth) distribution; Chen distribution; Type-II Heavy-Tailed Weibull (T2HTW) distribution; Generalized Exponential (GExp) distribution; Power Lindley (PLin) and Weibull (W) distribution; and Power Rayleigh (PRay) distribution.

8.1. Data Set I

The dataset consists of the follow-up times (in months) for 35 children undergoing growth hormone therapy, measured from the initiation of treatment until reaching the targeted developmental age: 2.15, 2.20, 2.55, 2.56, 2.63, 2.74, 2.81, 2.90, 3.05, 3.41, 3.43, 3.43, 3.84, 4.16, 4.18, 4.36, 4.42, 4.51, 4.60, 4.61, 4.75, 5.03, 5.10, 5.44, 5.90, 5.96, 6.77, 7.82, 8.00, 8.16, 8.21, 8.72, 10.40, 13.20, 13.70. Data were obtained from the Minas Gerais Health Secretariat Hormonal Program, Brazil, and were previously used in a published study to apply new statistical distributions in biomedical and epidemiological contexts [50].
Table 9 presents parameter and entropy estimates under different removal scenarios. With complete data, the results are automatically identical. It is observed that removal from the beginning or middle of the data affects the estimates to a greater extent than removal from the end, particularly at higher deletion rates, as reflected in the variation in Shannon entropy values and the widening of confidence intervals for the parameter λ . These results suggest that the impact of data removal depends not only on the deletion rate but also on the position of the removed values, which should be taken into account when dealing with incomplete data.
Table 10 shows that the IG distribution achieves the lowest information criterion values (AIC = 153.96, BIC = 157.53) and the highest p-values for the goodness-of-fit tests (AD p = 0.987, KS p = 0.979), indicating a better fit to the data. It is followed by GomLin, while the Chen and PMuth distributions show the weakest performance. The information criteria and goodness-of-fit tests agree on the ranking of distributions without contradiction, with a clear gap between IG and its closest competitor.
Figure 6 provides a comprehensive comparison of all distributions in terms of probability density and cumulative distribution functions. Most curves converge around the lower values, but differences appear in the upper tail of the data, where the IG distribution maintains a better representation of the general behavior compared to distributions that deviate from the empirical histogram.
Figure 7 displays the parameter estimation for the IG distribution via MLE. The top plots show a convex logarithmic likelihood profile, while the bottom plots present the score function intersecting z e r o , confirming that the values chosen for the parameters μ and λ are the optimal estimates.
Figure 8 compares empirical probabilities against theoretical probabilities for 12 distributions. The IG distribution exhibits close alignment with the red diagonal line, whereas other distributions, such as Chen and Plin, show noticeable deviation. These results indicate that the IG distribution provides a better fit for the data compared to the other models.
Figure 9 compares the empirical cumulative distribution function (black stepped line) with the theoretical functions of various distributions. Good agreement is observed for the IG and GomLin distributions with the empirical data, while other distributions show deviation in the tail regions, supporting the quality of the IG distribution in representing the cumulative structure of the data.

8.2. Data Set II

This dataset represents the survival times (in days) of 44 patients with Head and Neck Cancer (HNC) who underwent a combined radiation therapy and chemotherapy treatment regimen. 0.1220, 0.2356, 0.2374, 0.2587, 0.3198, 0.3700, 0.4135, 0.4738, 0.5546, 0.5836, 0.6347, 0.6846, 0.7447, 0.7826, 0.8143, 0.8400, 0.9200, 0.9400, 1.1000, 1.1200, 1.1900, 1.2700, 1.3000, 1.3300, 1.4000, 1.4600, 1.5500, 1.5900, 1.7300, 1.7900, 1.9400, 1.9500, 2.0900, 2.4900, 2.8100, 3.1900, 3.3900, 4.3200, 4.6900, 5.1900, 6.3300, 7.2500, 8.1700, 17.7600, showing a classic right-skewed distribution common in survival analysis, where most events occur early but a long tail extends to the right. Such data are typically modeled using heavy-tailed probability distributions (like the ZLindley or generalized Weibull models mentioned in the titles) to accurately capture the risk of mortality over time and to assess the efficacy of the simultaneous treatment protocol.
Table 11 presents parameter and entropy estimates under different removal scenarios. With complete data, the results are automatically identical. It is evident that removal from the middle (S2) and beginning (S3) affects estimates to a greater extent than removal from the end (S1), particularly at higher deletion rates, as reflected in the variation in Shannon entropy values and the widening of confidence intervals for the parameter μ . This pattern is consistent with what was observed for the growth hormone data, confirming that the impact of data removal depends on the position of the removed values and not only on the deletion rate. This should be taken into account when dealing with incomplete data.
Table 12 summarizes the comparative performance of the candidate probability distributions fitted to Dataset II. Among the competing models, the IG distribution consistently provides the best fit, as evidenced by the smallest information criteria (AIC = 160.38 and BIC = 163.49) and the largest goodness-of-fit p-values (AD p = 0.87, CvM p = 0.84, and KS p = 0.93). These findings identify the IG distribution as the most suitable model for Dataset II. The GExp and Gam distributions rank next in terms of model adequacy, whereas the Chen and PMuth distributions provide comparatively poorer fits. All criteria and tests agree on the ranking of distributions without contradiction, reinforcing confidence in the adoption of IG for modeling.
To avoid digression, Figure 10, Figure 11, Figure 12 and Figure 13 for dataset II also demonstrate the clear superiority of the Gauss inverse, as shown in Table 12 where results of dataset II confirm the pattern observed in dataset I, where the IG distribution is the best fit. All graphical displays (PP, CDF, Likelihood) showed consistent performance, in contrast to the clear poor fit of the other distributions.

9. Discussion

The simulation results, under the conditions considered in this study, showed that the MCMC and IS methods outperformed the MLE and Lindley methods in estimating both Shannon and Rényi entropies of the IG distribution. They achieved lower bias, smaller mean relative error, and more accurate confidence intervals. The performance of the MLE and its associated Wald confidence intervals deteriorated when the parameter values changed, particularly for Shannon entropy, whereas the MCMC and IS methods maintained relatively stable performance. The results also indicated, within the scope of this study, that Rényi entropy was less sensitive to changes in the parameter values than Shannon entropy, and that the GE and LINEX loss functions were more efficient than the SE loss function. Furthermore, the sensitivity analysis confirmed, for the settings examined, the robustness of the Bayesian estimates across the prior distributions considered.

10. Conclusions

The significance of this study is twofold:
  • Scientifically, it provides the first comprehensive systematic comparison of four estimation methods for Shannon and Rényi entropies of the IG distribution under progressive censoring, a topic largely unexplored in the literature. The results reveal that estimator stability depends fundamentally on the mathematical form of the entropy, with the Rényi entropy showing to be substantially less sensitive to parameter variation, while the Shannon entropy deteriorated markedly when the λ / μ ratio changed. The study further documents the breakdown of maximum likelihood estimation and Wald confidence intervals under complex scenarios, as well as the numerical instability of Lindley’s approximation at small sample sizes.
  • Practically, the study offers clear guidance in recommending MCMC and importance sampling as the optimal methods, with the general entropy and LINEX loss functions demonstrating superior performance. The stability of Bayesian estimates across different prior specifications further provides practitioners with flexibility even in the absence of precise prior information.

11. Future Research Directions

The findings of this study open several promising research avenues. First, the proposed methodological framework can be extended to other distributions with similar characteristics, such as generalized gamma distributions, to explore the performance of the different methods across broader distributional families. Second, extending the simulation framework to randomly distributed removal schemes, rather than single-stage withdrawals, would reveal whether our findings generalize to more realistic censoring scenarios.

Supplementary Materials

The following supporting information can be downloaded at: https://www.mdpi.com/article/10.3390/e28080871/s1.

Author Contributions

M.A.T.E.-S.: methodology (equal); writing—original draft (equal); writing—review and editing (equal); software (equal); formal analysis (equal); data curation (equal); R.A.H.: methodology (equal); writing—original draft (equal); writing—review and editing (equal); software (equal); formal analysis (equal); data curation (equal); D.B.: methodology (equal); writing—original draft (equal); writing—review and editing (equal); software (equal); formal analysis (equal); data curation (equal); T.A.: methodology (equal); writing—original draft (equal); writing—review and editing (equal); software (equal); formal analysis (equal); data curation (equal); W.S.A.E.A.: methodology (equal); writing—original draft (equal); writing—review and editing (equal); software (equal); formal analysis (equal); data curation (equal). All authors have read and agreed to the published version of the manuscript.

Funding

Princess Nourah bint Abdulrahman University Researchers Supporting Project number (PNURSP2026R404), Princess Nourah bint Abdulrahman University, Riyadh, Saudi Arabia.

Data Availability Statement

The data are available in the paper.

Acknowledgments

Princess Nourah bint Abdulrahman University Researchers Supporting Project number (PNURSP2026R404), Princess Nourah bint Abdulrahman University, Riyadh, Saudi Arabia.

Conflicts of Interest

The authors declare that they have no conflicts of interest.

Abbreviations

The following abbreviations are used in this manuscript:
ALAverage lengths.
CIsConfidence intervals.
CPCoverage probabilities.
EpWEpanechnikov–Weibull distribution.
GamGamma distribution.
GEGeneral Entropy loss function.
GExpGeneralized Exponential distribution.
GomLinGompertz-Lindley distribution.
HPDHighest posterior density.
ISImportance sampling.
MLsMaximum likelihood estimators
MREMean relative error.
MSEMean squared error.
MCMCMonte Carlo Markov Chain.
PerkPerks distribution.
PLinPower Lindley distribution.
PMuthPower Muth distribution.
PRayPower Rayleigh distribution.
RMSERoot mean squared error.
ScScheme.
SESquared Error loss function.
T2HTWType-II Heavy-Tailed Weibull distribution.
WWeibull distribution.
K ( . ) The Bessel function.
E 1 ( . ) The exponential integral.

References

  1. Shannon, C.E. A Mathematical Theory of Communication. Bell Syst. Tech. J. 1948, 27, 379–423. [Google Scholar] [CrossRef] [Scilit]
  2. Rényi, A. On measures of entropy and information. In Proceedings of the Fourth Berkeley Symposium on Mathematical Statistics and Probability; University of California Press: Oakland, CA, USA, 1961; Volume 4, pp. 547–562. [Google Scholar]
  3. Rass, S.; König, S. Password security as a game of entropies. Entropy 2018, 20, 312. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  4. Eun, S.-K.; Jung, N.-S.; Lee, J.-J.; Bae, Y.-J. Uncertainty of agricultural product prices by information entropy model using probability distribution for monthly prices. J. Korean Soc. Agric. Eng. 2012, 54, 7–14. [Google Scholar] [CrossRef] [Scilit]
  5. Harte, D.; Vere-Jones, D. The entropy score and its uses in earthquake forecasting. Pure Appl. Geophys. 2005, 162, 1229–1253. [Google Scholar] [CrossRef] [Scilit]
  6. Kamavaram, S.; Goseva-Popstojanova, K. Entropy as a measure of uncertainty in software reliability. In Proceedings of the 13th International Symposium on Software Reliability Engineering, Annapolis, MD, USA, 12–15 November 2002; pp. 209–210. [Google Scholar]
  7. Tweedie, M.C.K. A Mathematical Investigation of Some Electrophoretic Measurements on Colloids. Master’s Thesis, University of Reading, Reading, UK, 1941. [Google Scholar]
  8. Wald, A. On cumulative sums of random variables. Ann. Math. Stat. 1944, 15, 283–296. [Google Scholar] [CrossRef] [Scilit]
  9. Tweedie, M.C.K. Some statistical properties of inverse Gaussian distributions. Va. J. Sci. 1956, 7, 160–165. [Google Scholar] [CrossRef] [Scilit]
  10. Tweedie, M.C.K. Statistical properties of inverse Gaussian distributions–I. Ann. Math. Stat. 1957, 28, 362–377. [Google Scholar] [CrossRef] [Scilit]
  11. Tweedie, M.C.K. Statistical properties of inverse Gaussian distributions–II. Ann. Math. Stat. 1957, 28, 696–705. [Google Scholar] [CrossRef] [Scilit]
  12. Nath, G.B. Estimation of the parameter of the inverse Gaussian distribution from generalized censored samples. Metrika 1977, 24, 1–6. [Google Scholar] [CrossRef] [Scilit]
  13. Whitmore, G.A. A regression method for censored inverse-Gaussian data. Can. J. Stat. 1983, 11, 305–315. [Google Scholar] [CrossRef] [Scilit]
  14. Anaya, K.; O’Reilly, F. Test of fit for the inverse Gaussian and gamma distributions under censoring. Commun. Stat.–Theory Methods 2001, 30, 757–773. [Google Scholar] [CrossRef] [Scilit]
  15. Ismail, S.A.; Auda, H.A. Bayesian and fiducial inference for the inverse Gaussian distribution via Gibbs sampler. J. Appl. Stat. 2006, 33, 787–805. [Google Scholar] [CrossRef] [Scilit]
  16. Jia, J.M.; Yan, Z.Z.; Peng, X.Y. Estimation for inverse Gaussian distribution under first-failure progressive hybrid censored samples. Filomat 2017, 31, 5743–5752. [Google Scholar] [CrossRef] [Scilit]
  17. Rostamian, S.; Nematollahi, N. Estimation of stress–strength reliability in the inverse Gaussian distribution under progressively Type-II censored data. Math. Sci. 2019, 13, 175–191. [Google Scholar] [CrossRef] [Scilit]
  18. Jayalath, K.P.; Chhikara, R.S. Survival analysis for the inverse Gaussian distribution with the Gibbs sampler. J. Appl. Stat. 2022, 49, 656–675. [Google Scholar] [PubMed]
  19. Roy, S.; Pradhan, B.; Purakayastha, A. On inference and design under progressive type-I interval censoring scheme for inverse Gaussian lifetime model. Int. J. Qual. Reliab. Manag. 2022, 39, 1937–1962. [Google Scholar]
  20. Bera, S.; Jana, N. Estimating reliability parameters for inverse Gaussian distributions under complete and progressively Type-II censored samples. Qual. Technol. Quant. Manag. 2023, 20, 334–359. [Google Scholar]
  21. Xavier, T.; Vaisakh, K.M.; Sreedevi, E.P. Goodness of fit tests for inverse Gaussian distribution in presence and absence of censoring. J. Stat. Comput. Simul. 2025, 95, 1010–1032. [Google Scholar] [CrossRef] [Scilit]
  22. Kang, S.B.; Cho, Y.S.; Han, J.T.; Kim, J. An estimation of the entropy for a double exponential distribution based on multiply Type-II censored samples. Entropy 2012, 14, 161–173. [Google Scholar] [CrossRef] [Scilit]
  23. Cho, Y.; Sun, H.; Lee, K. An estimation of the entropy for a Rayleigh distribution based on doubly-generalized Type-II hybrid censored samples. Entropy 2014, 16, 3655–3669. [Google Scholar] [CrossRef] [Scilit]
  24. Cho, Y.; Sun, H.; Lee, K. Estimating the entropy of a Weibull distribution under generalized progressive hybrid censoring. Entropy 2015, 17, 102–122. [Google Scholar] [CrossRef] [Scilit]
  25. Anis, M.Z.; De, D. An expository note on Unit–Gompertz distribution with applications. Statistica 2020, 80, 469–490. [Google Scholar]
  26. Hashem, A.F.; Alyami, S.A. Inference on a new lifetime distribution under progressive Type-II censoring for a parallel-series structure. Complexity 2021, 1, 6684918. [Google Scholar]
  27. Okasha, H.; Nassar, M. Product of spacing estimation of entropy for inverse Weibull distribution under progressive Type-II censored data with applications. J. Taibah Univ. Sci. 2022, 16, 259–269. [Google Scholar] [CrossRef] [Scilit]
  28. Shrahili, M.; El-Saeed, A.R.; Hassan, A.S.; Elbatal, I.; Elgarhy, M. Estimation of entropy for log-logistic distribution under progressive Type-II censoring. J. Nanomater. 2022, 1, 2739606. [Google Scholar] [CrossRef] [Scilit]
  29. Maiti, K.; Kayal, S.; Kundu, D. Statistical inference on the Shannon and Rényi entropy measures of generalized exponential distribution under progressive censoring. SN Comput. Sci. 2022, 3, 1–21. [Google Scholar] [CrossRef] [Scilit]
  30. Patra, L.K.; Bajpai, S.; Misra, N. Inadmissibility of invariant estimator of function of scale parameter of several exponential distributions. arXiv 2023, arXiv:2302.2034. [Google Scholar]
  31. Ahmed, E.A.; El-Morshedy, M.; Al-Essa, L.A.; Eliwa, M.S. Statistical inference on the entropy measures of gamma distribution under progressive censoring: EM and MCMC algorithms. Mathematics 2023, 11, 2298. [Google Scholar] [CrossRef] [Scilit]
  32. Gong, Q.; Yin, B. Statistical inference of entropy functions of generalized inverse exponential model under progressive Type-II censoring test. PLoS ONE 2024, 19, e0311129. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  33. Folks, J.L.; Chhikara, R.S. The inverse Gaussian distribution and its statistical application—A review. J. R. Stat. Soc. Ser. B 1978, 40, 263–289. [Google Scholar] [CrossRef] [Scilit]
  34. Giner, G.; Smyth, G.K. statmod: Probability calculations for the inverse Gaussian distribution. R J. 2016, 8, 339–351. [Google Scholar] [CrossRef] [Scilit]
  35. Greene, W.H. Econometric Analysis, 4th ed.; Pearson Education India: New Delhi, India, 2003. [Google Scholar]
  36. Brown, L.D.; Cai, T.T.; DasGupta, A. Interval estimation for a binomial proportion. Stat. Sci. 2001, 16, 101–133. [Google Scholar] [CrossRef] [Scilit]
  37. Pandey, B.N.; Bandyopadhyay, P. Bayesian estimation of inverse Gaussian distribution. arXiv 2012, arXiv:1210.4524. [Google Scholar]
  38. Lindley, D.V. Approximate Bayes methods. In Bayesian Statistics; University of Valencia Press: Valencia, Spain, 1980. [Google Scholar]
  39. Nassar, M.M.; Eissa, F.H. Bayesian estimation for the exponentiated Weibull model. Commun. Stat.–Theory Methods 2005, 33, 2343–2362. [Google Scholar] [CrossRef] [Scilit]
  40. Kundu, D.; Pradhan, B. Bayesian inference and life testing plans for generalized exponential distribution. Sci. China Ser. A Math. 2009, 52, 1373–1388. [Google Scholar] [CrossRef] [Scilit]
  41. Xu, A.; Tang, Y. Reference analysis for Birnbaum–Saunders distribution. Comput. Stat. Data Anal. 2010, 54, 185–192. [Google Scholar] [CrossRef] [Scilit]
  42. Kim, C.; Jung, J.; Chung, Y. Bayesian estimation for the exponentiated Weibull model under Type II progressive censoring. Stat. Pap. 2011, 52, 53–70. [Google Scholar] [CrossRef] [Scilit]
  43. Sultan, H.; Ahmad, S.P. Bayesian approximation of generalized gamma distribution under three loss functions using Lindley’s approximation. J. Appl. Probab. Stat. 2015, 10, 21–33. [Google Scholar]
  44. Sana, S.; Faizan, M. Bayesian estimation using Lindley’s approximation and prediction of generalized exponential distribution based on lower record values. J. Stat. Appl. Probab. 2021, 10, 61–75. [Google Scholar] [CrossRef] [Scilit]
  45. Ren, H.; Hu, X. Bayesian estimations of Shannon entropy and Rényi entropy of inverse Weibull distribution. Mathematics 2023, 11, 2483. [Google Scholar] [CrossRef] [Scilit]
  46. Chen, M.H.; Shao, Q.M. Monte Carlo estimation of Bayesian credible and HPD intervals. J. Comput. Graph. Stat. 1999, 8, 69–92. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  47. Hastings, W.K. Monte Carlo sampling methods using Markov chains and their applications. Biometrika 1970, 57, 97–109. [Google Scholar] [CrossRef]
  48. Geman, S.; Geman, D. Stochastic relaxation, Gibbs distributions, and the Bayesian restoration of images. IEEE Trans. Pattern Anal. Mach. Intell. 1984, PAMI-6, 721–741. [Google Scholar] [CrossRef] [Scilit] [PubMed]
  49. Balakrishnan, N.; Sandhu, R.A. A simple simulational algorithm for generating progressive Type-II censored samples. Am. Stat. 1995, 49, 229–230. [Google Scholar] [CrossRef] [Scilit]
  50. Alizadeh, M.; Bagheri, S.F.; Samani, E.B.; Ghobadi, S.; Nadarajah, S. Exponentiated power Lindley power series class of distributions: Theory and applications. Commun. Stat.-Simul. Comput. 2018, 47, 2499–2531. [Google Scholar] [CrossRef] [Scilit]
Figure 1. Inverse Gaussian: (a) probability density function (PDF), (b) cumulative density function (CDF), (c) survival function (SRF), (d) hazard rate function (HRF).
Figure 1. Inverse Gaussian: (a) probability density function (PDF), (b) cumulative density function (CDF), (c) survival function (SRF), (d) hazard rate function (HRF).
Entropy 28 00871 g001
Figure 2. IG MCMC diagnostics density.
Figure 2. IG MCMC diagnostics density.
Entropy 28 00871 g002
Figure 3. IG MCMC diagnostics grid.
Figure 3. IG MCMC diagnostics grid.
Entropy 28 00871 g003
Figure 4. IG MCMC diagnostics density n50 m40 R1.
Figure 4. IG MCMC diagnostics density n50 m40 R1.
Entropy 28 00871 g004
Figure 5. IG MCMC diagnostics grid n50 m40 R1.
Figure 5. IG MCMC diagnostics grid n50 m40 R1.
Entropy 28 00871 g005
Figure 6. Fitted probability distributions for Dataset I.
Figure 6. Fitted probability distributions for Dataset I.
Entropy 28 00871 g006
Figure 7. Inverse Gaussian distribution diagnostics for Dataset I: log-likelihood profiles and score functions for both parameters.
Figure 7. Inverse Gaussian distribution diagnostics for Dataset I: log-likelihood profiles and score functions for both parameters.
Entropy 28 00871 g007
Figure 8. Probability– Probability (PP) plots for Dataset I under fitted candidate models.
Figure 8. Probability– Probability (PP) plots for Dataset I under fitted candidate models.
Entropy 28 00871 g008
Figure 9. Empirical versus fitted CDF plots for Dataset I under fitted candidate models.
Figure 9. Empirical versus fitted CDF plots for Dataset I under fitted candidate models.
Entropy 28 00871 g009
Figure 10. Fitted probability distributions for Dataset II.
Figure 10. Fitted probability distributions for Dataset II.
Entropy 28 00871 g010
Figure 11. Inverse Gaussian distribution diagnostics for Dataset II: log-likelihood profiles and score functions for both parameters.
Figure 11. Inverse Gaussian distribution diagnostics for Dataset II: log-likelihood profiles and score functions for both parameters.
Entropy 28 00871 g011
Figure 12. Probability–Probability (PP) plots for Dataset II under fitted candidate models.
Figure 12. Probability–Probability (PP) plots for Dataset II under fitted candidate models.
Entropy 28 00871 g012
Figure 13. Empirical versus fitted CDF plots for Dataset II under fitted candidate models.
Figure 13. Empirical versus fitted CDF plots for Dataset II under fitted candidate models.
Entropy 28 00871 g013
Table 1. Bayesian estimators under different loss functions.
Table 1. Bayesian estimators under different loss functions.
Loss FunctionBayes Estimator
S E ( H , H ^ ) = ( H ^ H ) 2 H ^ B . S E = E ( H | x ̲ )
G E ( H , H ^ ) = H ^ H q q ln H ^ H 1 , q 0 H ^ B . G E = E H q | x ̲ 1 q
L I N E X ( H , H ^ ) = e p ( H ^ H ) p ( H ^ H ) 1 , p 0 H ^ B . L I N E X = 1 p ln E e H p | x ̲
q > 0 , a positive error is more dangerous than a negative error. q < 0 , a negative error is more dangerous than a positive error.
Table 2. Progressive censoring schemes considered in the simulation study.
Table 2. Progressive censoring schemes considered in the simulation study.
SchemeRemoval Vector R                  Description
S1 R I = ( 0 , 0 , , 0 , n m ) R i = 0 for i = 1 , , m 1 , and R m = n m . All removals occur after the last observed failure.
S2 R II = ( 0 , , 0 , n m , 0 , , 0 ) R i = 0 for i m / 2 , and R m / 2 = n m . All removals occur at the middle failure time.
S3 R III = ( n m , 0 , 0 , , 0 ) R 1 = n m , and R i = 0 for i = 2 , , m . All removals occur after the first observed failure.
For all schemes, the condition i = 1 m R i = n m is satisfied.
Table 3. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Shannon entropy at Setting I ( μ = 1.5 , λ = 2 ).
Table 3. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Shannon entropy at Setting I ( μ = 1.5 , λ = 2 ).
nmSc.MetricPr. MLELindleyISMCMC
SE GE LINEX SE GE LINEX SE GE LINEX
3020S1Mean1.58501.87942.00811.92721.26501.21521.23551.26451.21441.2350
Bias0.33690.63130.75990.67900.0169−0.0329−0.01260.0163−0.0337−0.0132
MSE0.28290.46841.24930.72530.04080.04510.04050.04060.04540.0403
RMSE0.53190.68441.11770.85170.20190.21240.20120.20140.21300.2008
MRE0.33790.50830.61380.54720.12930.13620.12930.12890.13580.1288
3020S2Mean1.58311.81281.82781.80531.25791.21391.23191.26071.21631.2343
Bias0.33490.56470.57970.55720.0098−0.0342−0.01630.0125−0.0318−0.0139
MSE0.26000.39940.46100.39790.04380.04830.04390.04390.04820.0440
RMSE0.50990.63200.67900.63080.20930.21980.20960.20950.21960.2097
MRE0.32630.45500.46960.45000.13490.14180.13520.13510.14160.1352
3020S3Mean1.60131.77791.75131.75161.25421.21351.23031.25581.21471.2315
Bias0.35320.52970.50320.50340.0061−0.0346−0.01780.0076−0.0335−0.0166
MSE0.23700.36290.34890.33850.04430.04900.04450.04440.04890.0445
RMSE0.48680.60240.59060.58180.21040.22140.21100.21060.22120.2110
MRE0.32740.43250.41660.41350.13230.13800.13240.13300.13800.1324
5040S1Mean1.64661.75801.72381.72911.27131.24651.25581.27011.24581.2549
Bias0.39850.50990.47570.48090.0231−0.00160.00770.0220−0.00240.0068
MSE0.22120.30170.26490.26870.02120.02120.02060.02110.02110.0205
RMSE0.47030.54930.51470.51830.14580.14560.14340.14530.14530.1431
MRE0.32760.40850.38110.38530.09490.09400.09270.09490.09420.0930
5040S2Mean1.65871.74641.71451.71991.26961.24751.25581.26931.24751.2557
Bias0.41060.49830.46640.47180.0214−0.00060.00770.0212−0.00060.0076
MSE0.21980.28750.25560.25930.02020.02030.01970.02050.02050.0200
RMSE0.46880.53620.50560.50930.14220.14240.14040.14330.14330.1414
MRE0.33060.39920.37370.37800.09020.09080.08930.09140.09160.0903
5040S3Mean1.64181.72331.69511.70001.25621.23541.24341.25571.23511.2430
Bias0.39370.47520.44700.45190.0081−0.0127−0.00470.0076−0.0131−0.0051
MSE0.21130.26990.24270.24590.02310.02380.02290.02310.02380.0229
RMSE0.45970.51950.49270.49580.15210.15430.15150.15210.15430.1514
MRE0.32340.38300.36040.36430.09600.09650.09490.09600.09660.0950
10060S1Mean1.61071.71241.68101.68641.24871.22851.23631.24981.22911.2370
Bias0.36260.46430.43290.43830.0005−0.0197−0.01180.0017−0.0190−0.0111
MSE0.19650.26260.23060.23450.02290.02360.02270.02320.02390.0230
RMSE0.44330.51240.48020.48420.15130.15350.15080.15230.15440.1517
MRE0.31120.37450.34930.35340.09650.09640.09540.09640.09700.0956
10060S2Mean1.61781.68651.66221.66681.24301.22671.23301.24251.22631.2326
Bias0.36960.43830.41400.4186−0.0051−0.0214−0.0151−0.0056−0.0218−0.0155
MSE0.17900.22700.20540.20850.01770.01840.01780.01780.01850.0179
RMSE0.42300.47640.45320.45660.13300.13550.13330.13330.13600.1336
MRE0.30140.35210.33280.33640.08470.08540.08440.08490.08560.0845
10060S3Mean1.65561.70441.68391.68751.25461.24071.24601.25511.24111.2464
Bias0.40750.45630.43580.43930.0065−0.0075−0.00220.0069−0.0070−0.0017
MSE0.19890.23620.21760.22010.01450.01470.01430.01430.01450.0142
RMSE0.44590.48610.46650.46920.12020.12110.11970.11960.12040.1191
MRE0.32890.36590.34960.35240.07600.07640.07560.07570.07610.0753
10075S1Mean1.65601.71341.69001.69391.25551.24121.24661.25661.24221.2476
Bias0.40780.46530.44190.44580.0074−0.0069−0.00150.0085−0.0060−0.0005
MSE0.20690.25000.22790.23060.01610.01620.01590.01630.01640.0161
RMSE0.45480.50000.47740.48020.12670.12740.12610.12770.12810.1268
MRE0.32730.37280.35400.35720.08310.08310.08230.08430.08400.0834
10075S2Mean1.65071.69601.67691.68031.24771.23531.24001.24731.23511.2398
Bias0.40260.44790.42880.4321−0.0004−0.0128−0.0081−0.0008−0.0131−0.0084
MSE0.19380.22830.21130.21360.01400.01430.01400.01410.01440.0141
RMSE0.44020.47780.45960.46220.11840.11970.11830.11860.12000.1186
MRE0.32450.35950.34430.34690.07480.07570.07480.07520.07610.0752
10075S3Mean1.65501.69421.67751.68031.24831.23711.24141.24911.23791.2422
Bias0.40680.44600.42930.43220.0002−0.0110−0.00680.0010−0.0102−0.0059
MSE0.19930.22880.21390.21590.01500.01520.01490.01490.01510.0148
RMSE0.44650.47840.46250.46470.12240.12350.12230.12200.12300.1218
MRE0.32730.35780.34450.34670.07900.07990.07900.07900.07980.0790
True value: Shannon entropy = 1.2481.
Table 4. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Rényi entropy estimates ( α = 0.5 ) at Setting I ( μ = 1.5 , λ = 2 ).
Table 4. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Rényi entropy estimates ( α = 0.5 ) at Setting I ( μ = 1.5 , λ = 2 ).
nmSc.MetricPr. MLELindleyISMCMC
SE GE LINEX SE GE LINEX SE GE LINEX
3020S1Mean1.60261.85531.91261.87751.73231.67381.68341.73191.67341.6829
Bias−0.08680.16590.22320.18800.0429−0.0156−0.00600.0425−0.0160−0.0065
MSE0.13480.08460.28100.15110.06570.06590.06170.06550.06580.0617
RMSE0.36720.29090.53010.38870.25630.25680.24850.25600.25660.2483
MRE0.17420.14050.17900.15670.12070.12200.11790.12060.12160.1178
3020S2Mean1.60261.79821.79441.78741.71941.66901.67741.72291.67191.6803
Bias−0.08680.10870.10500.09800.0300−0.0204−0.01200.0335−0.0175−0.0091
MSE0.12010.07650.08740.07600.06960.07130.06730.06980.07130.0673
RMSE0.34660.27660.29560.27570.26390.26710.25940.26420.26690.2594
MRE0.16570.13300.13720.13010.12510.12680.12320.12540.12690.1232
3020S3Mean1.61911.76651.73951.74281.71361.66831.67601.71571.66981.6776
Bias−0.07030.07710.05010.05340.0242−0.0211−0.01340.0263−0.0196−0.0119
MSE0.09350.07330.07400.07020.06530.06750.06380.06570.06760.0639
RMSE0.30580.27070.27210.26490.25550.25970.25250.25620.26000.2528
MRE0.14170.12710.12530.12260.11920.12020.11710.11990.12030.1172
5040S1Mean1.65981.75271.72641.73051.73011.70021.70441.72871.69911.7033
Bias−0.02960.06330.03700.04110.04070.01080.01500.03930.00970.0139
MSE0.04760.03650.03210.03160.03560.03400.03310.03540.03390.0330
RMSE0.21820.19110.17910.17780.18860.18430.18190.18810.18410.1817
MRE0.10390.09130.08520.08470.09050.08810.08700.09060.08850.0873
5040S2Mean1.67001.74281.71821.72221.72761.70151.70531.72731.70151.7052
Bias−0.01940.05330.02880.03280.03820.01210.01590.03790.01210.0158
MSE0.03920.03370.03110.03060.03250.03130.03060.03300.03180.0311
RMSE0.19810.18370.17650.17490.18020.17690.17490.18160.17820.1762
MRE0.09670.08650.08330.08250.08560.08470.08370.08640.08560.0846
5040S3Mean1.65551.72261.70061.70431.70961.68571.68941.70891.68521.6888
Bias−0.03390.03320.01120.01490.0202−0.0037−0.00000.0195−0.0042−0.0006
MSE0.04420.03610.03450.03380.03600.03580.03490.03590.03570.0348
RMSE0.21030.19000.18580.18400.18960.18920.18680.18950.18900.1866
MRE0.09740.08890.08630.08560.08860.08790.08680.08830.08780.0867
10060S1Mean1.62931.71531.69171.69561.69611.67131.67521.69771.67221.6761
Bias−0.06010.02590.00230.00620.0067−0.0182−0.01430.0083−0.0172−0.0134
MSE0.05330.03690.03390.03350.03740.03740.03640.03810.03800.0369
RMSE0.23080.19220.18410.18300.19330.19330.19080.19510.19480.1922
MRE0.10700.09090.08710.08650.09120.09030.08930.09130.09110.0899
10060S2Mean1.63591.69361.67501.67841.68641.66701.67011.68561.66631.6694
Bias−0.05350.0042−0.0144−0.0110−0.0030−0.0224−0.0193−0.0038−0.0231−0.0200
MSE0.03560.02730.02700.02650.02780.02820.02760.02790.02840.0278
RMSE0.18860.16530.16430.16260.16660.16800.16610.16710.16860.1666
MRE0.08800.07860.07720.07660.07890.07850.07770.07900.07890.0780
10060S3Mean1.66791.70811.69231.69491.70361.68751.69001.70421.68801.6905
Bias−0.02150.01870.00290.00550.0142−0.00190.00060.0148−0.00140.0011
MSE0.02570.02240.02190.02150.02260.02240.02200.02230.02210.0218
RMSE0.16020.14970.14790.14680.15020.14960.14840.14930.14870.1475
MRE0.07450.07030.06910.06860.07030.06990.06940.07000.06970.0691
10075S1Mean1.66691.71481.69721.70011.70791.69041.69301.70951.69161.6942
Bias−0.02250.02540.00780.01070.01850.00100.00360.02010.00220.0048
MSE0.03120.02660.02550.02510.02630.02580.02540.02670.02620.0258
RMSE0.17670.16300.15960.15840.16210.16080.15950.16350.16180.1605
MRE0.08580.07960.07780.07720.07890.07810.07750.08030.07940.0788
10075S2Mean1.66241.70011.68561.68801.69661.68181.68411.69601.68141.6837
Bias−0.02700.0107−0.0038−0.00140.0072−0.0076−0.00530.0066−0.0080−0.0057
MSE0.02510.02180.02150.02120.02190.02190.02150.02200.02200.0216
RMSE0.15840.14760.14670.14560.14800.14790.14670.14820.14820.1470
MRE0.07440.06910.06870.06820.06930.06920.06860.06970.06970.0691
10075S3Mean1.66681.69891.68611.68831.69601.68301.68501.69701.68401.6860
Bias−0.02260.0095−0.0033−0.00110.0066−0.0064−0.00440.0076−0.0054−0.0034
MSE0.02640.02340.02320.02290.02370.02360.02330.02350.02350.0231
RMSE0.16250.15290.15220.15120.15380.15370.15270.15330.15310.1521
MRE0.07790.07310.07300.07250.07350.07370.07320.07360.07370.0732
True value: Rényi entropy ( α = 0.5 ) = 1.6894.
Table 5. Various confidence intervals of entropies measured under Setting I ( μ = 1.5 , λ = 2 ).
Table 5. Various confidence intervals of entropies measured under Setting I ( μ = 1.5 , λ = 2 ).
nmSc.EntropyMLE (Wald)IS (Credible)MCMC (CI)MCMC (HPD)
CP AL CP AL CP AL CP AL
3020S1Shannon1.0004.33000.9800.97270.9750.96320.9600.9487
Rényi1.0003.56850.9851.26580.9801.25610.9701.2279
3020S2Shannon1.0003.50370.9650.92050.9700.91570.9700.8999
Rényi1.0002.95310.9701.18010.9651.17630.9701.1482
3020S3Shannon1.0003.15410.9550.87090.9500.87500.9350.8618
Rényi1.0002.68490.9701.10300.9501.10920.9601.0845
5040S1Shannon1.0002.48270.9950.69770.9950.69090.9900.6798
Rényi1.0002.11590.9950.90660.9950.89820.9950.8785
5040S2Shannon1.0002.25280.9850.65990.9800.65110.9750.6431
Rényi1.0001.93530.9900.84600.9850.83660.9850.8210
5040S3Shannon0.9852.10810.9650.63220.9550.62900.9600.6215
Rényi1.0001.81970.9600.79900.9550.79460.9600.7815
10060S1Shannon1.0002.27000.9300.60790.9500.63240.9450.6227
Rényi1.0001.94100.9300.79580.9450.82780.9400.8109
10060S2Shannon0.9901.86720.9700.56000.9650.55560.9650.5483
Rényi1.0001.61670.9650.71640.9650.71100.9750.6995
10060S3Shannon0.9601.68730.9650.51870.9600.51930.9600.5134
Rényi1.0001.46570.9650.65480.9600.65480.9600.6449
10075S1Shannon0.9801.82720.9650.52630.9650.52930.9700.5227
Rényi1.0001.56880.9650.68390.9700.69070.9700.6792
10075S2Shannon0.9551.61560.9600.49060.9650.48420.9500.4797
Rényi1.0001.39870.9700.62910.9700.62070.9700.6128
10075S3Shannon0.8951.50210.9450.46310.9500.46250.9450.4576
Rényi1.0001.30550.9500.58450.9500.58340.9500.5765
Table 6. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Shannon entropy at Setting II ( μ = 2 , λ = 1 ).
Table 6. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Shannon entropy at Setting II ( μ = 2 , λ = 1 ).
nmSc.MetricPr. MLELindleyISMCMC
SE GE LINEX SE GE LINEX SE GE LINEX
3020S1Mean2.36502.06823.55312.77401.54091.49451.50931.54741.49651.5123
Bias0.80090.50401.98901.2098−0.0232−0.0696−0.0548−0.0167−0.0676−0.0518
MSE1.158522.4250100.01581.96970.05250.06430.05660.04290.05620.0483
RMSE1.07644.735510.00081.40340.22920.25360.23790.20710.23720.2198
MRE0.56211.18051.30090.78920.11050.11990.11370.10410.11450.1081
3020S2Mean2.30412.66122.66752.64191.54391.50481.51581.55051.51041.5211
Bias0.74001.09701.10341.0778−0.0203−0.0593−0.0483−0.0136−0.0537−0.0430
MSE0.80471.32051.37521.27660.03820.04480.04110.03540.04130.0387
RMSE0.89701.14911.17271.12990.19550.21170.20270.18810.20330.1967
MRE0.48920.70420.70860.69190.09830.10380.10000.09550.10210.0984
3020S3Mean2.34272.61792.57212.54561.55151.51381.52421.55091.51241.5230
Bias0.77851.05371.00800.9815−0.0126−0.0503−0.0400−0.0132−0.0518−0.0412
MSE0.81451.20781.12611.03890.03120.03690.03380.03010.03650.0330
RMSE0.90251.09901.06121.01930.17650.19220.18370.17340.19100.1817
MRE0.50510.67410.64470.62750.08770.09530.09140.08670.09480.0907
5040S1Mean2.41542.59512.54432.53051.56811.54331.54951.57691.54971.5563
Bias0.85121.03100.98020.96640.0039−0.0208−0.01470.0128−0.0144−0.0079
MSE0.90801.17261.03510.99930.02860.03000.02880.02650.02800.0268
RMSE0.95291.08291.01740.99960.16900.17320.16980.16270.16740.1638
MRE0.54450.66610.62670.61780.08820.09050.08860.08430.08700.0851
5040S2Mean2.35712.52192.47702.46581.54591.52381.52951.55101.52731.5333
Bias0.79300.95770.91290.9016−0.0182−0.0404−0.0346−0.0131−0.0368−0.0308
MSE0.76911.01280.91970.89450.03190.03440.03290.03050.03320.0316
RMSE0.87701.00640.95900.94580.17860.18550.18140.17450.18220.1779
MRE0.51170.61330.58360.57680.08960.09200.09030.08740.09040.0886
5040S3Mean2.38832.52232.47832.46801.55031.52851.53411.55111.52851.5342
Bias0.82420.95810.91420.9038−0.0138−0.0356−0.0301−0.0130−0.0356−0.0299
MSE0.82471.02150.93340.90880.03300.03530.03390.03200.03460.0331
RMSE0.90811.01070.96610.95330.18160.18790.18400.17890.18600.1819
MRE0.52930.61280.58480.57810.09320.09580.09410.09200.09520.0934
10060S1Mean2.43312.57742.52002.50371.55121.53271.53761.56871.54361.5497
Bias0.86901.01320.95580.9396−0.0130−0.0315−0.02660.0046−0.0205−0.0144
MSE0.98481.15521.01070.96920.04170.04360.04220.03150.03340.0320
RMSE0.99241.07481.00530.98450.20410.20870.20540.17760.18270.1788
MRE0.56030.64900.61110.60070.10470.10560.10420.09340.09450.0929
10060S2Mean2.37162.49232.45322.44461.54081.52521.52911.54891.53081.5352
Bias0.80750.92820.88910.8805−0.0233−0.0390−0.0350−0.0152−0.0333−0.0289
MSE0.75140.93770.86260.84340.02370.02490.02420.02140.02290.0221
RMSE0.86680.96830.92880.91840.15390.15790.15550.14640.15140.1487
MRE0.51690.59340.56840.56290.07880.08100.07980.07400.07670.0753
10060S3Mean2.41612.50582.47442.46701.56531.55031.55381.56751.55201.5556
Bias0.85200.94160.91020.90290.0012−0.0138−0.01030.0034−0.0122−0.0086
MSE0.81100.95540.89500.87890.02100.02150.02100.02070.02120.0207
RMSE0.90050.97740.94600.93750.14480.14670.14500.14380.14580.1440
MRE0.54470.60200.58190.57730.07660.07730.07650.07610.07690.0761
10075S1Mean2.46882.58042.53632.52431.58351.56821.57171.59001.57211.5761
Bias0.90461.01620.97210.96020.01930.00410.00760.02580.00800.0120
MSE0.94091.11171.01740.98970.02530.02530.02490.02280.02280.0224
RMSE0.97001.05441.00870.99480.15920.15900.15770.15110.15110.1496
MRE0.57830.64970.62150.61390.07970.07970.07900.07540.07540.0746
10075S2Mean2.40422.49072.46062.45381.55101.53801.54111.55331.53911.5425
Bias0.84000.92650.89640.8897−0.0132−0.0262−0.0230−0.0108−0.0250−0.0216
MSE0.77730.91690.86030.84600.01880.01960.01910.01830.01920.0187
RMSE0.88160.95750.92750.91980.13720.14000.13830.13530.13860.1368
MRE0.53700.59240.57310.56880.06950.07080.06990.06900.07070.0698
10075S3Mean2.43862.51032.48452.47831.57131.55921.56201.57421.56151.5643
Bias0.87440.94620.92040.91420.0072−0.0049−0.00220.0101−0.00260.0002
MSE0.83330.95340.90380.89050.01770.01790.01760.01760.01780.0175
RMSE0.91290.97640.95070.94370.13320.13390.13280.13260.13330.1321
MRE0.55910.60490.58840.58450.06900.06880.06840.06820.06800.0676
True value: Shannon entropy = 1.5641. note: The numerical collapse of Shannon entropy values reflects the amplified MLE errors, a phenomenon explicitly noted by Giner and Smyth [34] who stated that the IG distribution is “extremely difficult to treat numerically” and that Newton–Raphson algorithm “cannot in general be guaranteed to converge”, particularly in the extreme tails. It is worth noting that Shannon entropy is highly sensitive to extreme values, making it more vulnerable to estimator instability compared to other entropy measures.
Table 7. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Rényi entropy estimates ( α = 0.5 ) at Setting II ( μ = 2 , λ = 1 ).
Table 7. Mean, Bias, MSE, RMSE, and MRE of ML and Bayes estimates for Rényi entropy estimates ( α = 0.5 ) at Setting II ( μ = 2 , λ = 1 ).
nmSc.MetricPr. MLELindleyISMCMC
SE GE LINEX SE GE LINEX SE GE LINEX
3020S1Mean2.21712.00412.56962.56262.30682.23792.23342.31632.23952.2343
Bias−0.0825−0.29560.27000.26290.0071−0.0618−0.06630.0166−0.0602−0.0653
MSE0.365613.20700.49800.47780.10630.11900.11270.08370.09930.0930
RMSE0.60473.63410.70570.69130.32600.34500.33580.28930.31510.3050
MRE0.20740.40630.20740.19670.10860.11270.10950.10090.10610.1028
3020S2Mean2.17392.45792.44342.43702.29662.23892.23392.30672.24632.2411
Bias−0.12580.15820.14370.1373−0.0030−0.0608−0.06580.0071−0.0534−0.0586
MSE0.19600.10860.10210.10080.07780.08610.08230.07180.08140.0778
RMSE0.44280.32960.31950.31750.27900.29340.28680.26790.28540.2788
MRE0.15440.12480.11470.11030.09760.09960.09720.09440.09830.0961
3020S3Mean2.20442.41872.37572.37162.31352.25902.25362.31162.25652.2511
Bias−0.09520.11910.07600.07200.0138−0.0407−0.04610.0119−0.0432−0.0485
MSE0.15520.08370.07360.07860.06660.07280.07000.06360.07110.0684
RMSE0.39400.28930.27130.28030.25810.26980.26460.25220.26660.2615
MRE0.13560.10670.09370.09090.08790.09060.08890.08570.08970.0883
5040S1Mean2.26492.40672.36832.36182.32872.28982.28482.34422.30052.2945
Bias−0.03480.10700.06860.06210.0290−0.0098−0.01480.04450.0008−0.0051
MSE0.12700.08630.05920.05370.06610.06610.06400.06140.06120.0591
RMSE0.35630.29370.24330.23170.25710.25700.25290.24790.24740.2430
MRE0.12600.10320.08960.08540.09110.09130.08980.08790.08740.0859
5040S2Mean2.21672.34602.31152.30632.29152.25842.25462.30042.26462.2603
Bias−0.08300.04630.01190.0066−0.0082−0.0413−0.04510.0007−0.0351−0.0393
MSE0.10680.07050.06380.06110.06530.06800.06630.06290.06600.0643
RMSE0.32680.26560.25250.24720.25550.26070.25750.25090.25680.2536
MRE0.11200.09280.08750.08540.08640.08790.08680.08480.08640.0853
5040S3Mean2.23952.34382.31032.30542.30802.27632.27232.30892.27592.2718
Bias−0.06020.04410.01060.00570.0083−0.0234−0.02730.0093−0.0238−0.0279
MSE0.10670.07550.07070.06760.07200.07340.07150.06860.07070.0690
RMSE0.32670.27480.26590.26010.26840.27100.26740.26200.26600.2626
MRE0.11460.09860.09500.09280.09410.09490.09360.09260.09370.0924
10060S1Mean2.27462.39122.34922.34142.31232.28122.27742.33852.29542.2895
Bias−0.02500.09160.04950.04170.0126−0.0185−0.02230.0388−0.0042−0.0102
MSE0.15780.09630.07210.06610.10260.10350.10080.07770.07710.0742
RMSE0.39730.31040.26860.25700.32040.32170.31760.27880.27760.2724
MRE0.13920.11650.10040.09550.11260.11120.10950.10150.09920.0971
10060S2Mean2.22572.32272.29352.28952.28362.25932.25622.29732.26852.2648
Bias−0.07400.0230−0.0062−0.0102−0.0161−0.0404−0.0434−0.0023−0.0311−0.0349
MSE0.07480.05360.05090.04900.05370.05490.05380.04920.05050.0495
RMSE0.27350.23140.22550.22140.23170.23430.23190.22170.22470.2224
MRE0.09460.08100.07880.07720.08000.08130.08060.07710.07780.0770
10060S3Mean2.26522.33532.31162.30812.31612.29392.29072.31962.29652.2931
Bias−0.03440.03570.01190.00840.0164−0.0058−0.00900.0199−0.0032−0.0066
MSE0.06060.04950.04720.04570.04670.04650.04560.04650.04620.0452
RMSE0.24610.22240.21720.21380.21620.21570.21350.21550.21490.2127
MRE0.09010.08110.07940.07820.07860.07860.07770.07870.07860.0778
10075S1Mean2.30682.39552.36312.35732.35282.32742.32342.36412.33362.3286
Bias0.00710.09580.06340.05770.05310.02770.02370.06440.03400.0290
MSE0.08380.06390.05510.05200.06240.06010.05840.05690.05410.0524
RMSE0.28950.25280.23470.22810.24970.24510.24170.23860.23260.2288
MRE0.09820.08860.08190.07960.08580.08430.08310.08160.07950.0782
10075S2Mean2.25152.32032.29792.29472.30252.28232.27952.30672.28442.2812
Bias−0.04820.0207−0.0018−0.00490.0028−0.0174−0.02010.0070−0.0153−0.0184
MSE0.05290.04180.04050.03930.04060.04090.04020.03950.03990.0392
RMSE0.23000.20440.20120.19840.20160.20220.20050.19880.19990.1981
MRE0.08060.07050.06970.06880.07020.07070.07010.06860.06950.0690
10075S3Mean2.28392.33972.32032.31732.32362.30572.30312.32832.30942.3065
Bias−0.01570.04000.02060.01760.02400.00610.00340.02870.00970.0068
MSE0.04800.04210.04020.03910.04000.03950.03880.04000.03930.0385
RMSE0.21900.20530.20050.19780.20010.19870.19690.19990.19810.1962
MRE0.07730.07270.07100.07000.07110.07040.06970.07110.07020.0695
True value: Rényi entropy ( α = 0.5 ) = 2.2997 .
Table 8. Various confidence intervals of entropies measured under Setting II ( μ = 2 , λ = 1 ).
Table 8. Various confidence intervals of entropies measured under Setting II ( μ = 2 , λ = 1 ).
nmSc.EntropyMLE (Wald)IS (Credible)MCMC (CI)MCMC (HPD)
CP AL CP AL CP AL CP AL
3020S1Shannon1.0009.07750.9150.92740.9901.01980.9901.0120
Rényi1.0006.90960.9201.43470.9951.59310.9951.5724
3020S2Shannon1.0005.28480.9800.91230.9950.94420.9950.9361
Rényi1.0004.28270.9851.37680.9901.43840.9901.4170
3020S3Shannon1.0004.72850.9850.89900.9950.92500.9950.9167
Rényi1.0003.87060.9801.36460.9951.38740.9951.3638
5040S1Shannon1.0004.31350.9700.72580.9950.79630.9950.7882
Rényi1.0003.48980.9551.12570.9901.25660.9901.2351
5040S2Shannon0.9553.36540.9400.70040.9500.73990.9500.7318
Rényi0.9952.77310.9751.05900.9651.13040.9651.1105
5040S3Shannon0.8953.19100.9400.70070.9600.72470.9550.7167
Rényi1.0002.63820.9401.04580.9551.08730.9551.0674
10060S1Shannon1.0004.51800.8350.58200.9750.76320.9650.7561
Rényi1.0003.63300.8300.95930.9801.24590.9601.2243
10060S2Shannon0.8952.91260.8850.55290.9750.65270.9750.6448
Rényi1.0002.40630.9100.86270.9751.01630.9750.9972
10060S3Shannon0.6702.55350.9600.58650.9650.60910.9600.6028
Rényi1.0002.12280.9600.88020.9650.91680.9750.9015
10075S1Shannon0.9303.26800.9050.55690.9600.65540.9450.6476
Rényi1.0002.66190.8850.89120.9701.06030.9501.0397
10075S2Shannon0.7252.48430.9300.52670.9600.58000.9450.5732
Rényi1.0002.05560.9400.80900.9750.89630.9750.8792
10075S3Shannon0.5002.28200.9550.53520.9700.55290.9600.5469
Rényi1.0001.89720.9650.79880.9650.82830.9600.8151
Table 9. Growth Hormone Therapy Data: Parameter, CI, Shannon entropy estimates and Rényi entropy estimates at α = 0.5 .
Table 9. Growth Hormone Therapy Data: Parameter, CI, Shannon entropy estimates and Rényi entropy estimates at α = 0.5 .
Schemem μ λ Shannon HRényi H α
Estimate 95% CI Estimate 95% CI
0% removalS1355.3057[4.403, 6.208]20.1311[10.699, 29.563]2.24432.5331
S2355.3057[4.403, 6.208]20.1311[10.699, 29.563]2.24432.5331
S3355.3057[4.403, 6.208]20.1311[10.699, 29.563]2.24432.5331
10% removalS1315.1840[4.298, 6.070]21.5931[10.749, 32.437]2.18882.4694
S2315.6278[4.555, 6.700]19.0071[9.806, 28.209]2.34202.6419
S3315.6823[4.735, 6.630]25.2381[12.980, 37.496]2.25772.5331
20% removalS1285.3686[4.325, 6.413]19.5151[9.130, 29.901]2.27062.5633
S2285.9031[4.666, 7.140]18.3109[9.234, 27.388]2.41682.7257
S3285.9890[4.993, 6.985]29.6103[14.640, 44.581]2.27102.5383
30% removalS1254.9556[4.088, 5.823]24.4811[10.666, 38.296]2.08192.3493
S2256.1370[4.710, 7.564]17.5208[8.645, 26.397]2.48112.7993
S3256.3098[5.234, 7.385]33.2309[15.560, 50.901]2.29942.5623
Note: CI = Confidence Interval (95%); Rényi entropy with α = 0.5 . S1: Removal at end; S2: Removal at middle; S3: Removal at beginning.
Table 10. Distribution comparison results for dataset I.
Table 10. Distribution comparison results for dataset I.
Dist.AICAICCBICHQICCAICAD StatAD pCvM StatCvM pKS StatKS p
IG153.9581154.2508157.5264155.2814159.52640.21200.98680.03610.95390.06790.9790
GomLin159.9374160.2301163.5058161.2607165.50580.75370.51480.10870.54560.10570.6704
T2HTW162.4722163.0722167.8248164.4572170.82480.67480.57930.10690.55390.11720.5420
PLin163.5180163.8106167.0863164.8413169.08630.94230.38850.15050.38940.13060.4061
W162.4283162.7210165.9967163.7516167.99670.87420.42990.14320.41230.13070.4053
PRay162.4283162.7210165.9967163.7516167.99670.87420.42980.14320.41230.13070.4052
EpW163.1794163.4721166.7478164.5028168.74780.92270.39990.14840.39570.13070.4048
Perk162.7648163.0575166.3332164.0882168.33320.96440.37600.17690.31790.14520.2836
Gam162.7488163.0415166.3172164.0722168.31720.98720.36360.18480.29970.14730.2680
PMuth166.5134166.8061170.0818167.8368172.08181.27600.24020.21050.24840.14730.2680
GExp162.6551162.9478166.2235163.9784168.22351.02460.34420.19710.27370.14980.2508
Chen176.6565176.9492180.2249177.9798182.22492.25090.06740.36520.08910.18310.0918
Table 11. Radiation Therapy Data: Parameter, CI, Shannon entropy estimates and Rényi entropy estimates at α = 0.5 .
Table 11. Radiation Therapy Data: Parameter, CI, Shannon entropy estimates and Rényi entropy estimates at α = 0.5 .
Schemem μ λ Shannon HRényi H α
Estimate 95% CI Estimate 95% CI
0% removalS1442.2348[1.321, 3.148]1.1678[0.680, 1.656]1.67952.3978
S2442.2348[1.321, 3.148]1.1678[0.680, 1.656]1.67952.3978
S3442.2348[1.321, 3.148]1.1678[0.680, 1.656]1.67952.3978
10% removalS1402.1771[1.174, 3.180]1.1832[0.669, 1.697]1.65682.3601
S2402.4982[1.331, 3.666]1.1504[0.664, 1.637]1.77692.5456
S3402.4321[1.495, 3.370]1.5718[0.885, 2.259]1.77742.4183
20% removalS1352.0782[1.001, 3.155]1.2091[0.657, 1.761]1.61522.2932
S2352.8619[1.266, 4.458]1.1272[0.645, 1.609]1.88902.7248
S3352.7110[1.643, 3.779]1.9181[1.021, 2.815]1.88782.4984
30% removalS1312.0098[0.867, 3.153]1.2298[0.644, 1.816]1.58452.2444
S2313.2307[1.124, 5.337]1.1133[0.632, 1.594]1.98422.8814
S3312.9690[1.714, 4.224]2.0582[1.036, 3.081]1.97852.5958
Note: CI = Confidence Interval (95%); Rényi entropy with α = 0.5 . S1: Removal at end; S2: Removal at middle; S3: Removal at beginning.
Table 12. Distribution comparison results for dataset II.
Table 12. Distribution comparison results for dataset II.
Dist.AICAICCBICHQICCAICAD StatAD pCvM StatCvM pKS StatKS p
IG160.3819160.7569163.4926161.4557165.49260.37740.87000.05710.83540.09150.9315
GExp162.1951162.5701165.3058163.2689167.30580.52600.71930.07650.71540.10120.8657
Gam164.2165164.5915167.3272165.2903169.32720.66330.58910.10450.56580.12350.6594
T2HTW167.5767168.3509172.2428169.1874175.24280.74880.51840.11730.50890.13410.5545
Perk170.9981171.3731174.1088172.0719176.10880.96180.37730.12370.48260.13780.5196
GomLin170.9932171.3682174.1039172.0670176.10390.96140.37750.12360.48320.13790.5187
PLin167.7614168.1364170.8721168.8353172.87210.87770.42740.13790.43040.14080.4916
W168.9773169.3523172.0879170.0511174.08790.97700.36900.15180.38550.14500.4536
PRay168.9772169.3522172.0879170.0510174.08790.97890.36790.15250.38350.14540.4500
EpW169.6924170.0674172.8031170.7662174.80311.01810.34730.15700.37030.14580.4465
PMuth171.7277172.1027174.8384172.8015176.83841.23260.25520.19610.27580.15730.3516
Chen178.8560179.2310181.9667179.9298183.96671.75780.12560.27600.15810.19200.1515
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content.

Share and Cite

MDPI and ACS Style

El-Shahat, M.A.T.; Hashem, R.A.; Basalamah, D.; Alballa, T.; Azm, W.S.A.E. A Computational Bayesian Framework for Entropy-Based Inference in the Inverse Gaussian Distribution Under Progressive Type-II Censoring. Entropy 2026, 28, 871. https://doi.org/10.3390/e28080871

AMA Style

El-Shahat MAT, Hashem RA, Basalamah D, Alballa T, Azm WSAE. A Computational Bayesian Framework for Entropy-Based Inference in the Inverse Gaussian Distribution Under Progressive Type-II Censoring. Entropy. 2026; 28(8):871. https://doi.org/10.3390/e28080871

Chicago/Turabian Style

El-Shahat, Mohamed A. T., Reman Abo Hashem, Doaa Basalamah, Tmader Alballa, and Wael S. Abu El Azm. 2026. "A Computational Bayesian Framework for Entropy-Based Inference in the Inverse Gaussian Distribution Under Progressive Type-II Censoring" Entropy 28, no. 8: 871. https://doi.org/10.3390/e28080871

APA Style

El-Shahat, M. A. T., Hashem, R. A., Basalamah, D., Alballa, T., & Azm, W. S. A. E. (2026). A Computational Bayesian Framework for Entropy-Based Inference in the Inverse Gaussian Distribution Under Progressive Type-II Censoring. Entropy, 28(8), 871. https://doi.org/10.3390/e28080871

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop