← all papers · overview

Efedllm: Efficient LLM Inference Based On Federated Learning

Abstract

Large Language Models (LLMs) herald a transformative era in artificial intelligence (AI). However, the expansive scale of data and parameters of LLMs requires high-demand computational and memory resources, restricting their accessibility to a broader range of users and researchers. This paper introduces an effective approach that enhances the operational efficiency and affordability of LLM infere

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).