Barely Significant
← all excerpts

Assessing Artificial Intelligence (AI) in Patient Education: Evaluating Accuracy and Readability of Responses on Surgical Procedures for Patellar Tendon Rupture.

Cureus · 2026 · PMC13286298 · PMID 42338855

2
hedged sentences
0.0580
closest p · 1.2× alpha
0.0750
boldest claim

The sentences

did not reach statistical significancep=0.058so close (0.05 < p ≤ 0.1)
DISCERN score Perplexity demonstrated the highest mean DISCERN score among the evaluated models; however, the overall Kruskal-Wallis test did not reach statistical significance (p=0.058).

also in 111,027 other papers

approached significancep = 0.075so close (0.05 < p ≤ 0.1)
Even though Perplexity demonstrated the highest mean DISCERN scores among the evaluated AI models, no statistically significant differences in readability were observed among the four chatbots, although results approached significance (p = 0.075).

also in 8,237 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.