← all papers · overview

Robust Implementation Of Retrieval-augmented Generation On Edge-based Computing-in-memory Architectures

Abstract

Large Language Models (LLMs) deployed on edge devices learn through fine-tuning and updating a certain portion of their parameters. Although such learning methods can be optimized to reduce resource utilization, the overall required resources remain a heavy burden on edge devices. Instead, Retrieval-Augmented Generation (RAG), a resource-efficient LLM learning method, can improve the quality of th

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).