Robustness Over Time: Understanding Adversarial Examples' Effectiveness on Longitudinal Versions of Large Language Models
时间鲁棒性:对抗示例在大型语言模型纵向版本中的有效性
专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract)
AI总结 本文研究了大型语言模型在纵向版本中的对抗鲁棒性,发现更新并不总是提升鲁棒性,部分模型在误分类和幻觉方面有所下降。