arXiv cs.LG· Narek Maloyan·· 3 小时前AI 评分29
评估并提升大语言模型对输入序列变化的鲁棒性
Evaluating and Improving the Robustness of Large Language Models to Input Sequence Variations
AI 导读
一篇博士论文提出 R_stab(f) 生成式鲁棒性度量,并开发 ASA 自适应进化黑盒攻击,对 LLM-as-a-Judge 系统攻击成功率最高达 73.8%,开源模型间迁移率最高 62.6%。
来源:arXiv cs.LG · arxiv.org