This test resulted in a p -value of 0.025 and a positive trend of 0.302, statistically verifying that as we move from ‘Definitely AI’ to ‘Definitely Human’ on the scale, the proportion of human-authored submissions increases.
← all excerpts
A comparison of human, GPT-3.5, and GPT-4 performance in a university-level coding course.
1
—
—