In all cases, the agreement shows an increasing trend from F1 to F3, while for F4 the highest disagreement has been obtained due to the inability of ChatGPT to recognize the features related to this stage, as previously illustrated.
← all excerpts
Assessing the diagnostic accuracy of ChatGPT-4 in the histopathological evaluation of liver fibrosis in MASH.
1
—
—