← all papers · overview

Beyond Numeric Rewards: In-context Dueling Bandits With LLM Agents

Abstract

In-Context Reinforcement Learning (ICRL) is a frontier paradigm to solve Reinforcement Learning (RL) problems in the foundation model era. While ICRL capabilities have been demonstrated in transformers through task-specific training, the potential of Large Language Models (LLMs) out-of-the-box remains largely unexplored. This paper investigates whether LLMs can generalize cross-domain to perform I

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).