The data indicates a clear trend of improvement: as the model size grows, the number of correct answers increases noticeably, from 604 in the smallest model (8B) to 912 in the largest (405B).
← all excerpts
Generalist large language models in a specialized world: Evidence from the Italian national medical education pathway.
1
—
—