Reinforcement learning (RL) fine-tuning improves vision-language-action (VLA) policies through closed-loop experience, yet generalization beyond the fine-tun…
机构:南大
来源:arXiv 2610.09943 | AI4Papers 论文推荐平台