Barely Significant
← all excerpts

Large language model chatbots as sources of pediatric anesthesia health advice: An evaluation of reliability and readability.

Digit Health · 2026 · PMC13319765 · PMID 42389384

1
hedged sentence
closest p
boldest claim

The sentences

DeepSeek and Gemini demonstrated relatively higher median reliability scores overall; however, pairwise comparisons showed that differences between Claude and either DeepSeek or Gemini did not reach statistical significance.

also in 111,027 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.