Don't Judge a Book by its Cover: Testing LLMs' Robustness Under Logical Obfuscation
不要只看封面判断一本书:测试LLMs在逻辑混淆下的鲁棒性
机构 * Manipal University Jaipur(马纳普大学贾普尔分校) ; Indian Institute of Technology, Patna(印度理工学院帕坦分校) ; Indian Institute of Science Education and Research, Kolkata(印度科学教育与研究学院科钦分校)
专题命中 具身导航 :navigation(abstract)
AI总结 本研究提出Logifus框架和LogiQAte基准,测试LLMs在逻辑混淆下的鲁棒性,发现混淆显著降低模型性能,揭示当前LLMs缺乏深度理解。
Comments 19 pages, 6 figures