机构
*
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of California, Merced(加州大学默塞德分校)
;
Southeast University(东南大学)
Governed Individuation: Cryptographically Decoupling an Agent's Learning from Its Authority
受治理的个体化:通过密码学方式将智能体的学习与其授权方解耦
Xue Qin, Simin Luan, Cong Yang, Zhijun Li
机构
*
School of Software, Harbin Institute of Technology(哈尔滨工业大学软件学院)
;
School of Computer Science and Technology, Harbin Institute of Technology(哈尔滨工业大学计算机科学与技术学院)
;
School of Future Science and Engineering, Soochow University(苏州大学未来科学与工程学院)
机构
*
Computer Network Information Center, Chinese Academy of Sciences(中国科学院计算机网络信息中心)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
University of Manchester(曼彻斯特大学)
;
Institute of Metal Research, Chinese Academy of Sciences(中国科学院金属研究所)
;
China University of Geosciences(中国地质大学)
机构
*
Research Center for Social Computing and Interactive Robotics(社会计算与交互机器人研究中心)
;
Department of Computer Science and Technology, Institute for AI(计算机科学与技术系,人工智能研究院)
机构
*
Shenzhen Loop Area Institute(深圳河套学院)
;
Dalian University of Technology(大连理工大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems
超越古德哈特定律:多智能体系统中合规性评估的动态基准
Yiyang Zhao, Zhuo Zhang, Qingxuan Le, Lizhen Qu, Zenglin Xu
机构
*
Fudan University(复旦大学)
;
Shanghai Academy of AI for Science(上海人工智能科学研究院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Monash University(莫纳什大学)
CommentsPeer-reviewed and presented at the 1st Workshop on Toward Trustworthy Vision-Language Models in the Wild (TrustVLM), co-located with ACM ICMR 2026, Amsterdam. Non-archival workshop. Reviews public on OpenReview. 5 pages, 2 figures
Humans Are More Diverse: Frontier LLMs Show Extreme Policies in Idealised AI Development Races
人类更多样化:前沿大语言模型(LLM)在理想化AI开发竞赛中表现出极端策略
Phu Hoa Pham, Duy Minh Dao Sy, Trung Kiet Huynh, Phu Quy Nguyen Lam, Chi Nguyen Tran, Minh Trung Le, Phong Hao Le, Dinh Nam Nguyen, Thien Ky Nguyen Dong, Elias Fernandez Domingos, Le Hong Trang, The Anh Han
机构
*
Ho Chi Minh City University of Science(胡志明市科学大学)
;
Vietnam National University Ho Chi Minh City (VNU-HCM)(越南国家大学胡志明市分校)
;
Ho Chi Minh City University of Technology (HCMUT)(胡志明市技术大学(HCMUT))
Unifying Adversarially Robust Model Experts in Vision-Language Models
统一视觉-语言模型中的对抗鲁棒模型专家
Nguyen Duc Thai, Junhao Dong, Sua Qi Rong, Hua Yu, Yew-Soon Ong
机构
*
College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
;
Center for Frontier AI Research, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心)