GPT-4 performed poorer than ChatGPT in assigning differential diagnosis (22.7%), but this did not reach statistical significance ( P = 0.13).
← all excerpts
Comparison of ChatGPT-3.5 and GPT-4 as potential tools in artificial intelligence-assisted clinical practice in renal and liver transplantation.
1
0.1300
0.1300