← all papers · overview

Mindllm: Pre-training Lightweight Large Language Model From Scratch, Evaluations And Domain Applications

Abstract

Large Language Models (LLMs) have demonstrated remarkable performance across various natural language tasks, marking significant strides towards general artificial intelligence. While general artificial intelligence is leveraged by developing increasingly large-scale models, there could be another branch to develop lightweight custom models that better serve certain domains, taking into account th

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).