Risk aversion in resources could prevent misaligned AI agents from causing catastrophic harm. Misaligned but risk-averse agents would tend to favor safer str…
机构:Columbia University
来源:arXiv 2609.38093 | AI4Papers 论文推荐平台