In this paper, we propose trajectory advantage regression, a method of offline path learning and path attribution based on reinforcement learning. The proposed method can be used to solve path optimization problems while algorithmically only solving a regression problem.
Related papers
Ranked by semantic similarity β how closely each paper's abstract matches this one (100% = near-identical topic).