Barely Significant
← all excerpts

Comparative Analysis of ChatGPT-4o and Gemini Advanced Performance on Diagnostic Radiology In-Training Exams.

Cureus · 2025 · PMC12009162 · PMID 40255788

1
hedged sentence
0.0500
closest p · 1.0× alpha
0.0500
boldest claim

The sentences

did not reach statistical significancep > 0.05actually significant
Statistical testing across categories showed that differences in accuracy between the two models did not reach statistical significance (p > 0.05 for all comparisons).

also in 111,027 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.