← all papers · overview

Can Small Language Models Use What They Retrieve? An Empirical Study Of Retrieval Utilization Across Model Scale

Abstract

Retrieval augmented generation RAG is widely deployed to improve factual accuracy in language models yet it remains unclear whether smaller models of size 7B parameters or less can effectively utilize retrieved information. To investigate this question we evaluate five model sizes from 360M to 8B across three architecture families SmolLM2 Qwen2.5 and Llama 3.1 under four retrieval conditions inclu

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).