← all papers · overview

Multilinguality In Llm-designed Reward Functions For Restless Bandits: Effects On Task Performance And Fairness

Abstract

Restless Multi-Armed Bandits (RMABs) have been successfully applied to resource allocation problems in a variety of settings, including public health. With the rapid development of powerful large language models (LLMs), they are increasingly used to design reward functions to better match human preferences. Recent work has shown that LLMs can be used to tailor automated allocation decisions to com

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).