← all papers · overview

POET-X: Memory-efficient LLM Training By Scaling Orthogonal Transformation

Abstract

Efficient and stable training of large language models (LLMs) remains a core challenge in modern machine learning systems. To address this challenge, Reparameterized Orthogonal Equivalence Training (POET), a spectrum-preserving framework that optimizes each weight matrix through orthogonal equivalence transformation, has been proposed. Although POET provides strong training stability, its original

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).