Barely Significant
← all excerpts

Comparative Performance of ChatGPT 3.5 and GPT4 on Rhinology Standardized Board Examination Questions.

OTO Open · 2024 · PMC11208739 · PMID 38938507

2
hedged sentences
0.0001
closest p · 0.0× alpha
0.0001
boldest claim

The sentences

highly significantP < .0001actually significant
The comparison between ChatGPT 3.5's score (45.16%) and the average resident score (76.34%) yielded a highly significant difference, with P < .0001.

also in 132,142 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.