Vision-language models (VLMs) have demonstrated remarkable zero-shot capabilities yet remain sensitive to real-world distribution shifts during inference. Al…
机构:UIUC
来源:arXiv 2608.18339 | AI4Papers 论文推荐平台