← all papers · overview

Mediswift: Efficient Sparse Pre-trained Biomedical Language Models

Abstract

Large language models (LLMs) are typically trained on general source data for various domains, but a recent surge in domain-specific LLMs has shown their potential to outperform general-purpose models in domain-specific tasks (e.g., biomedicine). Although domain-specific pre-training enhances efficiency and leads to smaller models, the computational costs of training these LLMs remain high, posing

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).