The results show a clear trend where moderate chunk sizes (e.g., 300 tokens) and no overlap produced the best outcomes, balancing context preservation and retrieval granularity.
← all excerpts
Scalable evaluation framework for retrieval augmented generation in tobacco research using large Language models.
1
—
—