While GPT-4o demonstrated higher accuracy than GPT-4 across the image-based items, this difference did not reach statistical significance, likely due to the limited sample size.
← all excerpts
The performance of ChatGPT on medical image-based assessments and implications for medical education.
1
—
—