Barely Significant
← all excerpts

Accuracy Is Not Enough: Reasoning and Reference Reliability in Orthopaedic Large Language Model (LLM) Applications.

Cureus · 2026 · PMC12874175 · PMID 41658703

1
hedged sentence
0.5200
closest p · 10.4× alpha
0.5200
boldest claim

The sentences

Image-based questions demonstrated lower accuracy (44.7%, 17/38) compared with text-based questions (54%, 27/50), though this difference did not reach statistical significance (Fisher's exact test, no conventional test statistic; p=0.52) (Table 3 ).

also in 111,027 other papers

Quoted from the open-access full text in Europe PMC under the licence the publisher applied. The sentence is reproduced exactly as published; the emphasis is ours.