Barely Significant
← all excerpts

Comparison of ChatGPT-3.5 and GPT-4 as potential tools in artificial intelligence-assisted clinical practice in renal and liver transplantation.

World J Transplant · 2025 · PMC12038595 · PMID 40881761

1
hedged sentence
0.1300
closest p · 2.6× alpha
0.1300
boldest claim

The sentences

did not reach statistical significanceP = 0.13not close (p > 0.1)
GPT-4 performed poorer than ChatGPT in assigning differential diagnosis (22.7%), but this did not reach statistical significance ( P = 0.13).

also in 111,027 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.