As multi-agent systems enter high-stakes domains, the possibility that agents may circumvent safety boundaries is a growing concern. Prior work has examined …
机构:UIUC
来源:arXiv 2609.39050 | AI4Papers 论文推荐平台