(2) Sequence length dependence: In the control variable experiments, when the input sequence lengths are 2k, 4k and 6k tokens, the model’s retrieval accuracy for the first document shows a decreasing trend (77% → 75% → 72%), indicating that long sequence processing exacerbates the model performance decay.
1
—
—