Barely Significant
← all excerpts

Performance comparison of large language models for medication counseling in people living with HIV.

Front Public Health · 2026 · PMC13260627 · PMID 42293577

1
hedged sentence
0.0010
closest p · 0.0× alpha
0.0010
boldest claim

The sentences

highly significantp < 0.001actually significant
Results The comprehensive scores ranked from highest to lowest were DeepSeek (4.47), Qwen (4.33), Kimi (4.24), Doubao (4.13), and ChatGPT (3.41), with highly significant differences were observed among all models ( H =182.14, p < 0.001).

also in 132,142 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.