Barely Significant
← all excerpts

Performance evaluation of five major large language models in tuberculosis Q&A systems: A multidimensional assessment of readability, quality, and reliability.

Digit Health · 2026 · PMC13354909 · PMID 42436893

1
hedged sentence
0.0010
closest p · 0.0× alpha
0.0010
boldest claim

The sentences

highly significantP<0.001actually significant
These differences were highly significant (F=32.21, P<0.001), indicating that texts generated by GPT-5 performed best in terms of understandability and actionability as patient education materials.

also in 132,142 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.