At the other extreme, one must be cautious in the interpretation of calibration results with small cohorts, because, even when the calibration curve, the calibration intercept and slope points to a miscalibration, the p values of traditional calibration statistics may not be significant, raising concern about the study low power [ 25 ].
← all excerpts
External validation of SAPS 3 and MPM<sub>0</sub>-III scores in 48,816 patients from 72 Brazilian ICUs.
2
—
—
The sentences
There is a known phenomenon with traditional calibration statistics (such as Hosmer–Lemeshow goodness of fit) in prediction models validation/calibration studies with many thousands included subjects, in which often p values are highly significant despite visually good calibration curves, very small absolute errors, and acceptable calibration slope and intercept.