highly significantp < 0.001
Pairwise comparisons revealed that the performance difference between XAI-enhanced and traditional ML approaches was not statistically significant ( p = 0.23, paired t -test), while the interpretability improvement was highly significant ( p < 0.001, Wilcoxon signed-rank test).