机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Qatar Computing Research Institute(卡塔尔计算研究所)
;
Hamad Bin Khalifa University(哈马德·本·卡伊夫大学)
DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and its Loss' Convexity is Dispensable)
DPO 解绑:你的训练算法在人类选择理论中实际上是解耦的(且其损失函数的凸性并非必需)
Wenxuan Zhou, Shujian Zhang, Brice Magdalou, John Lambert, Ehsan Amid, Richard Nock, Andrew Hard
机构
*
Google Research(谷歌研究)
;
Google DeepMind(谷歌DeepMind)
;
Work done while at Google DeepMind(在谷歌深Mind的工作)
;
CEE-M, Univ Montpellier, CNRS, INRAE, Institut Agro(CEE-M,蒙彼利埃大学,CNRS,INRAE,农业研究所)
ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training
ALOE: 面向视觉-语言-动作模型后训练的动作级离策略评估
Rushuai Yang, Hecheng Wang, Zhichao Wu, Chiming Liu, Xiaohan Yan, Xuan Du, Shuoyu Yue, Chuheng Zhang, Yunlong Wang, Yongcheng Liu, Lizhe Qi, Yi Chen, Wei Shan, Maoqing Yao
机构
*
AgiBot
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Fudan University(复旦大学)
;
Nanjing University(南京大学)
;
Independent Researcher(独立研究者)
Positive Alignment: Artificial Intelligence for Human Flourishing
积极对齐:人工智能促进人类繁荣
Ruben Laukkonen, Seb Krier, Chloé Bakalar, Shamil Chandaria, Morten Kringelbach, Adam Elwood, Daniel Ford, Fernando Rosas, Maty Bohacek, Matija Franklin, Nenad Tomašev, Stephanie Chan, Verena Rieser, Roma Patel, Michael Levin, Arun Rao
机构
*
Department of Psychiatry, University of Oxford(牛津大学精神病学系)
;
Flourishing Intelligence Program, Centre for Eudaimonia and Human Flourishing, Linacre College, University of Oxford(牛津大学幸福智能计划、幸福与人类繁荣中心、林acre学院)
;
Google DeepMind(谷歌DeepMind)
;
LIFE
;
OpenAI
;
Anthropic
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Aily Labs(Aily实验室)
;
Stanford University(斯坦福大学)
;
Tufts University(塔夫茨大学)
;
Positive AI Labs(积极AI实验室)
;
Department of Informatics, University of Sussex(Sussex大学信息学系)
;
Department of Brain Sciences, Imperial College London(伦敦帝国理工学院脑科学系)
CommentsProceedings of the 39th Annual Conference on Neural Information Processing Systems, ARLET Workshop (Aligning Reinforcement Learning Experimentalists and Theorists)
Journal refTransactions on Machine Learning Research, Vol. 2026, June 2026
SPOT-E: Test-Time Entropy Shaping with Visual Spotlights for Frozen VLMs
SPOT-E:基于视觉聚光灯的冻结VLM测试时熵整形
Bo Yin, Xiaobin Hu, Chengming Xu, Ruolin Shen, Mo Yang, Jiangning Zhang, Peng-Tao Jiang, Cheng Tan, Shuicheng Yan
机构
*
National University of Singapore(新加坡国立大学)
;
Fudan University(复旦大学)
;
Technical University of Munich(慕尼黑工业大学)
;
Sagenic Tech
;
Zhejiang University(浙江大学)
;
vivo
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
Backdoor Channels Hidden in Latent Space: Cryptographic Undetectability in Modern Neural Networks
潜伏空间中的后门通道:现代神经网络中的密码学不可检测性
Marte Eggen, Eirik Reiestad, Kristian Gjøsteen, Inga Strümke
机构
*
Department of Computer Science, Norwegian University of Science and Technology(挪威科学技术大学计算机科学系)
;
Department of Mathematical Sciences, Norwegian University of Science and Technology(挪威科学技术大学数学科学系)
How Your Credentials Are Leaked by LLM Agent Skills: An Empirical Study
你的凭证如何被LLM智能体技能泄露:一项实证研究
Zhihao Chen, Ying Zhang, Yi Liu, Gelei Deng, Yuekang Li, Yanjun Zhang, Jianting Ning, Leo Yu Zhang, Lei Ma, Zhiqiang Li
机构
*
Griffith University(格里菲斯大学)
;
Wake Forest University(威克森林大学)
;
Nanyang Technological University(南洋理工大学)
;
University of New South Wales(新南威尔士大学)
;
Zhejiang Sci-Tech University(浙江科技学院)
;
The University of Tokyo(东京大学)
;
University of Alberta(阿尔伯塔大学)
;
Independent Researcher(独立研究者)
专题命中
长上下文与记忆
:LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI