Barely Significant
← all excerpts

Can large language models be trusted? Reliability and readability of responses to perinatal depression FAQs.

Front Public Health · 2026 · PMC12968175 · PMID 41810293

1
hedged sentence
0.0640
closest p · 1.3× alpha
0.0640
boldest claim

The sentences

did not reach statistical significancep = 0.064so close (0.05 < p ≤ 0.1)
The Friedman test results indicated statistically significant between model differences for ARI, CLI, OLWF, LWGLF, and FRF (all p < 0.01), whereas GFI did not reach statistical significance ( p = 0.064).

also in 111,027 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.