← all papers · overview

Openba-v2: Reaching 77.3% High Compression Ratio With Fast Multi-stage Pruning

Abstract

Large Language Models (LLMs) have played an important role in many fields due to their powerful capabilities.However, their massive number of parameters leads to high deployment requirements and incurs significant inference costs, which impedes their practical applications. Training smaller models is an effective way to address this problem. Therefore, we introduce OpenBA-V2, a 3.4B model derived

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).