It is noteworthy that interobserver agreement between the first and second sessions showed an increasing trend, which suggested that more thorough training with the scale (the training was reviewed before the second round of scoring) and increased user familiarity can improve the reliability of the scale.
1
—
—