机构
*
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
School of Information Science and Engineering, Wuhan University of Science and Technology(武汉科技大学信息科学与工程学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
Combining Adam and its Inverse Counterpart to Enhance Generalization of Deep Learning Optimizers
将Adam及其逆向对应物结合以增强深度学习优化器的泛化能力
Tao Shi, Liangming Chen, Long Jin, Mengchu Zhou
机构
*
School of Information Science and Engineering, Lanzhou University(兰州大学信息科学与工程学院)
;
Helen and John C. Hartmann Department of Electrical and Computer Engineering, New Jersey Institute of Technology(新泽西理工学院电气与计算机工程系)
专题命中
指令微调
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG
"Dark Triad" Model Organisms of Misalignment: Narrow Fine-Tuning Mirrors Human Antisocial Behavior
黑暗三联征模型生物:对齐偏差:狭窄微调映射人类反社会行为
Roshni Lulla, Fiona Collins, Sanaya Parekh, Thilo Hagendorff, Jonas Kaplan
机构
*
Brain & Creativity Institute, University of Southern California(大脑与创造力研究所,南加州大学)
;
Department of Psychology, University of Southern California(心理学系,南加州大学)
;
Interchange Forum for Reflecting on Intelligent Systems, University of Stuttgart(智能系统反思交流论坛,斯图加特大学)
专题命中
指令微调
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Stealth Fine-Tuning: Efficiently Breaking Alignment in RVLMs Using Self-Generated CoT
隐形微调:通过自动生成的CoT打破RVLMs的对齐
Le Yu, Zhengyue Zhao, Yawen Zheng, Yunhao Liu
机构
*
Machine Intelligence Laboratory, Sichuan University(四川大学人工智能实验室)
;
University of Wisconsin--Madison(威斯康星大学麦迪逊分校)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
Global Innovation Exchange, Tsinghua University(清华大学全球创新交流中心)
机构
*
Huawei Foundation Model Department(华为基础模型部门)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
专题命中
后训练与偏好优化
:language model(title,abstract);large language model(abstract);post-training(abstract);分类 cs.AI、cs.LG