borderline significantp = 0.05
FKGL of simplified texts of ChatGPT-3.5/-4.0 and Bard/Gemini was borderline significant (MD:-1.59; CI:-3.15,-0.04; p = 0.05) and subgroup-analysis between ChatGPT-4.0 and Bard was not significant (MD:-1.68; CI:-3.53,0.17; p = 0.07).