Towards Mitigation of Hallucination for LLM-empowered Agents: Progressive Generalization Bound Exploration and Watchdog Monitor
迈向减轻基于大语言模型的智能体幻觉:渐进泛化边界探索与看门狗监测
Siyuan Liu, Wenjing Liu, Zhiwei Xu, Xin Wang, Bo Chen, Tao Li
机构
*
College of Computer Science, Nankai University(南开大学计算机科学学院)
;
Haihe Lab of ITAI College of intelligent Science and Technology, Inner Mongolia University of Technology(内蒙古工业大学智能科学与技术学院海河实验室)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Department of Electrical and Computer Engineering, Stony Brook University(石溪大学电气与计算机工程系)
;
Department of Computer Science, Michigan Technological University(密歇根技术大学计算机科学系)
Comments14 pages, 2 figures, 6 tables. Preregistered on OSF (https://osf.io/feka7, DOI 10.17605/OSF.IO/FEKA7). Materials-availability and deviations described in the paper
CommentsAlso available at https://doi.org/10.5281/zenodo.20487041 (v6). v2: the pre-registered continuation-gate experiment has been run -- prediction (i) confirmed on real frontier agents; prediction (ii) retired to mathematical scope. Code, pre-registrations, and raw run records: https://github.com/macrokit/value
机构
*
Independent Researcher(独立研究者)
;
Hunan University(湖南大学)
;
Shenzhen Kaihong Digital Industry Development Co., Ltd.(深圳凯鸿数字产业开发有限公司)
;
Central University of Finance and Economics(中央财经大学)
;
Chongqing University(重庆大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Stevens Institute of Technology(斯蒂文斯理工学院)
;
Cornell University(康奈尔大学)
;
Oak Ridge National Laboratory(橡树岭国家实验室)
;
Zhengzhou University of Light Industry(郑州轻工业大学)
;
Xidian University(西安电子科技大学)
;
The University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
AI Safety Research Lab, Institute of Advanced Computing(高级计算研究所人工智能安全研究实验室)
;
University of Malaya(马来亚大学)
;
Emory University(埃默里大学)
;
Department of Nephrology, Affiliated Hospital of Guangdong Medical University(广东医学院附属医院肾内科)
Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models
使用潜在空间生成世界模型减轻自动驾驶模仿学习中的协变量转移
Alexander Popov, Alperen Degirmenci, David Wehr, Shashank Hegde, Ryan Oldja, Alexey Kamenev, Bertrand Douillard, David Nistér, Urs Muller, Ruchi Bhargava, Stan Birchfield, Nikolai Smolyanskiy
Comments8 pages, 6 figures, original September 2024, accepted at ICRA 2025 Workshop "Robots in the Wild", for associated video file, see https://youtu.be/7m3bXzlVQvU