highly significantp < 0.001
In contrast, DeepSeek received 70.39% Excellent, 22.37% Good, 7.02% Normal, and only 0.22% Awful rating, with a highly significant difference observed between the two models ( p < 0.001).
In contrast, DeepSeek received 70.39% Excellent, 22.37% Good, 7.02% Normal, and only 0.22% Awful rating, with a highly significant difference observed between the two models ( p < 0.001).