We also observed that DeepSeek achieved higher mean accuracy than clinicians in all subgroups, although the differences did not reach statistical significance, possibly due to the small number of clinicians included in each subgroup.
← all excerpts
Guideline adherence in surgical decisions for T1 colorectal cancer after endoscopic resection: large language models vs clinicians.
1
—
—