Similar results were found for correctness, although the effect did not reach statistical significance ( p = 0.059).
← all excerpts
Performance of ChatGPT in Pediatric Audiology as Rated by Students and Experts.
2
0.0590
0.0650
The sentences
almost reached significancep = 0.065
However, the differences between the ratings by students and by experts of question 15 were not statistically significant (although in terms of relevance, the difference almost reached significance, p = 0.065).