← all papers · overview

Tlora: Efficient Multi-lora Training With Elastic Shared Super-models

Abstract

As Low-Rank Adaptation (LoRA) becomes the standard approach for efficiently fine-tuning large language models (LLMs), shared clusters increasingly execute many concurrent LoRA training jobs over the same frozen backbone. While recent advances enable batching (co-locating) multiple adapters during serving, efficient training-time co-location of heterogeneous LoRA adapters presents unique challenges

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).