机构
*
Qwen Large Model Application Team, Alibaba(阿里巴巴大模型应用团队)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Shenzhen Research Institute of Big Data(深圳大数据研究院)
专题命中
后训练与偏好优化
:RLHF(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Contextualized Visual Personalization in Vision-Language Models
基于上下文的视觉个性化在视觉-语言模型中
Yeongtak Oh, Sangwon Yu, Junsung Park, Han Cheol Moon, Jisoo Mok, Sungroh Yoon
机构
*
Department of Electrical and Computer Engineering, Seoul National University, Seoul, South Korea(电气电子工程系,首尔国立大学,首尔,韩国)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University, Seoul, Korea(人工智能交叉学科项目,首尔国立大学,首尔,韩国)
机构
*
College of Computer, National University of Defense Technology(国防科技大学计算机学院)
;
Intelligent Game and Decision Lab (IGDL)(智能游戏与决策实验室)
;
Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究院)
专题命中
后训练与偏好优化
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
Noise-corrected GRPO: From Noisy Rewards to Unbiased Gradients
噪声校正的GRPO:从噪声奖励到无偏梯度
Omar El Mansouri, Fathinah Asma Izzati, Mohamed El Amine Seddik, Salem Lahlou
机构
*
Department of Machine Learning, Mohamed bin Zayed University of Artificial Intelligence, Abu Dhabi, UAE
;
Technology Innovation Institute, Abu Dhabi, UAE
;
Department of Robotics, Khalifa University, Abu Dhabi, UAE
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
RoVLA: 多一致性约束用于鲁棒的视觉-语言-动作模型
Jingzhou Luo, Yifan Wen, Yongjie Bai, Xinshuai Song, Yang Liu, Liang Lin
机构
*
Sun Yat-sen University(中山大学)
;
Peng Cheng Laboratory(鹏城实验室)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)
;
X-Era AI Lab(X-Era AI实验室)