Debate to Align: Reliable Entity Alignment through Two-Stage Multi-Agent Debate
辩论以对齐:通过双阶段多智能体辩论实现可靠的实体对齐
Cunda Wang, Ziying Ma, Po Hu, Weihua Wang, Feilong Bao
机构
*
Hubei Provincial Key Laboratory of Artificial Intelligence and Smart Learning, Central China Normal University, Wuhan, China(湖北人工智能与智能学习省级重点实验室,中央财经大学,武汉,中国)
;
School of Computer Science, Central China Normal University, Wuhan, China(中央财经大学计算机科学学院,武汉,中国)
;
National Language Resources Monitoring and Research Center for Network Media, Central China Normal University, Wuhan, China(网络媒体语言资源监测与研究中心,中央财经大学,武汉,中国)
;
College of Computer Science, Inner Mongolia University, Hohhot, China(内蒙古大学计算机学院,呼和浩特,中国)
;
National and Local Joint Engineering Research Center of Intelligent Information Processing Technology for Mongolian, Inner Mongolia University, Hohhot, China(蒙古语智能信息处理技术国家与地方联合工程研究中心,内蒙古大学,呼和浩特,中国)
;
Inner Mongolia Key Laboratory of Multilingual Artificial Intelligence Technology, Inner Mongolia University, Hohhot, China(内蒙古多语言人工智能技术重点实验室,内蒙古大学,呼和浩特,中国)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.CL
机构
*
Tsinghua University(清华大学)
;
Hefei University of Technology(合肥工业大学)
;
University of Arizona(亚利桑那大学)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统实验室)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);post-training(abstract);分类 cs.AI
FINER: MLLMs Hallucinate under Fine-grained Negative Queries
FINER:MLLMs在细粒度负查询下产生幻觉
Rui Xiao, Sanghwan Kim, Yongqin Xian, Zeynep Akata, Stephan Alaniz
机构
*
Technical University of Munich(慕尼黑技术大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
Helmholtz Munich(海德堡-慕尼黑亥姆霍尔茨中心)
;
Google(谷歌)
;
LTCI, Télécom Paris, Institut Polytechnique de Paris(LTCI,巴黎电信学院,巴黎理工学院)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);preference optimization(abstract);分类 cs.AI
Swadesh Jana, Cansu Sancaktar, Tomáš Daniš, Georg Martius, Antonio Orvieto, Pavel Kolev
机构
*
University of Tübingen(图宾根大学)
;
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
ELLIS Institute Tübingen(图宾根ELLIS研究所)
;
Tübingen AI Center(图宾根人工智能中心)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);post-training(abstract);分类 cs.LG
CommentsAccepted at ICLR 2026 Workshop on AI with Recursive Self-Improvement (RSI 2026) as Spotlight, and ICLR 2026 Workshop on Lifelong Agents (LLA 2026)
Low-Rank Contextual Reinforcement Learning from Heterogeneous Human Feedback
低秩上下文强化学习从异质人类反馈
Seong Jin Lee, Will Wei Sun, Yufeng Liu
机构
*
Department of Statistics and Operations Research, University of North Carolina, Chapel Hill(统计与运筹学系,北卡罗来纳大学 Chapel Hill 分校)
;
Daniels School of Business, Purdue University(商务学院,普渡大学)
;
Department of Statistics and Operations Research, Department of Genetics, Department of Biostatistics, The University of North Carolina at Chapel Hill(统计与运筹学系,遗传学系,生物统计学系,北卡罗来纳大学 Chapel Hill 分校)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);RLHF(abstract);分类 cs.LG
ExPO-HM: Learning to Explain-then-Detect for Hateful Meme Detection
ExPO-HM:为仇恨表情包检测学习解释-然后检测
Jingbiao Mei, Mingsheng Sun, Jinghong Chen, Pengda Qin, Yuhong Li, Da Chen, Bill Byrne
机构
*
Department of Engineering, University of Cambridge(剑桥大学工程系)
;
Xiaohongshu Inc.(小红书公司)
;
University of Bath(巴斯大学)
;
Tencent Company, China(腾讯公司)
;
Alibaba Group(阿里巴巴集团)
Zhuofan Josh Ying, Shauli Ravfogel, Nikolaus Kriegeskorte, Peter Hase
机构
*
Department of Psychology, Zuckerman Mind Brain Behavior Institute, Columbia University, New York, NY(心理学系、Zuckerman mind brain behavior研究所、哥伦比亚大学、纽约NY)
;
Department of Neuroscience, Columbia University, New York, NY(神经科学系、哥伦比亚大学、纽约NY)
;
New York University, New York, NY(纽约大学、纽约NY)
;
Stanford University, Stanford, CA(斯坦福大学、斯坦福CA)
专题命中
后训练与偏好优化
:large language model(abstract);language model(abstract);post-training(abstract);分类 cs.LG