← all papers · overview

Enhancing The General Agent Capabilities Of Low-parameter Llms Through Tuning And Multi-branch Reasoning

Abstract

Open-source pre-trained Large Language Models (LLMs) exhibit strong language understanding and generation capabilities, making them highly successful in a variety of tasks. However, when used as agents for dealing with complex problems in the real world, their performance is far inferior to large commercial models such as ChatGPT and GPT-4. As intelligent agents, LLMs need to have the capabilities

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).