← all papers · overview

Omnirouter: Budget And Performance Controllable Multi-llm Routing

Abstract

Large language models (LLMs) deliver superior performance but require substantial computational resources and operate with relatively low efficiency, while smaller models can efficiently handle simpler tasks with fewer resources. LLM routing is a crucial paradigm that dynamically selects the most suitable large language models from a pool of candidates to process diverse inputs, ensuring optimal r

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).