Reinforcement learning (RL) for reachability specifications is fundamental to sequential decision-making. Prior work establishes asymptotic convergence to op…
机构:Georgia Tech
来源:arXiv 2610.01781 | AI4Papers 论文推荐平台