When Do Language Models Endorse Limitations on Human Rights Principles?
语言模型何时会支持人权原则的限制?
Keenan Samway, Nicole Miu Takagi, Rada Mihalcea, Bernhard Schölkopf, Ilias Chalkidis, Daniel Hershcovich, Zhijing Jin
机构
*
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
Jinesis AI Lab, University of Toronto & Vector Institute(Jinesis AI实验室,多伦多大学及向量研究所)
;
University of Michigan(密歇根大学)
;
University of Copenhagen(哥本哈根大学)
;
EuroSafeAI
Low-Rank Contextual Reinforcement Learning from Heterogeneous Human Feedback
低秩上下文强化学习从异质人类反馈
Seong Jin Lee, Will Wei Sun, Yufeng Liu
机构
*
Department of Statistics and Operations Research, University of North Carolina, Chapel Hill(统计与运筹学系,北卡罗来纳大学 Chapel Hill 分校)
;
Daniels School of Business, Purdue University(商务学院,普渡大学)
;
Department of Statistics and Operations Research, Department of Genetics, Department of Biostatistics, The University of North Carolina at Chapel Hill(统计与运筹学系,遗传学系,生物统计学系,北卡罗来纳大学 Chapel Hill 分校)
Comments9 pages, 9 figures. Jake Thomas served as Editor for this manuscript
Journal refProceedings of the 2025 Conference on Applied Machine Learning for Information Security, Proceedings of Machine Learning Research PMLR Vol 299 pp 28 41
机构
*
School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院)
;
School of Biomedical Engineering, Shanghai Jiao Tong University(上海交通大学生物医学工程学院)
;
Longwood Valley MedTech
CommentsWithdrawn due to a critical error discovered in the stability and convergence proofs (specifically Lemma 2, Theorem 12, and Proposition 10) in Section 3. The identified flaw invalidates the core theoretical guarantees regarding capability growth and system stability
Cognition to Control - Multi-Agent Learning for Human-Humanoid Collaborative Transport
认知到控制 - 多智能体学习用于人-仿人协作运输
Hao Zhang, Ding Zhao, H. Eric Tseng
机构
*
Department of Electrical Engineering, the University of Texas at Arlington(电气工程系,德克萨斯大学阿灵顿分校)
;
Department of Mechanical Engineering, Carnegie Mellon University(机械工程系,卡内基梅隆大学)
Lihu Chen, Gerard de Melo, Fabian M. Suchanek, Gaël Varoquaux
机构
*
Imperial College London(伦敦帝国学院)
;
Hasso Plattner Institute / University of Potsdam(霍普夫纳研究所/波茨坦大学)
;
Telecom Paris, Institut Polytechnique de Paris(电信巴黎学院,巴黎理工学院)
;
Soda, Inria Saclay(Soda,法国国家信息与自动化技术研究院萨克利实验室)
From Privacy to Trust in the Agentic Era: A Taxonomy of Challenges in Trustworthy Federated Learning Through the Lens of Trust Report 2.0
从隐私到信任在代理时代:通过信任报告2.0的视角,对可信联邦学习中挑战的分类
Nuria Rodríguez-Barroso, Mario García-Márquez, M. Victoria Luzón, Francisco Herrera
机构
*
Department of Computer Science and Artificial Intelligence, Andalusian Research Institute in Data Science and Computational Intelligence (DaSCI) University of Granada(计算机科学与人工智能系,数据科学与计算智能安达卢西亚研究 institute,格拉纳达大学)
;
Department of Software Engineering, Andalusian Research Institute in Data Science and Computational Intelligence (DaSCI) University of Granada(软件工程系,数据科学与计算智能安达卢西亚研究 institute,格拉纳达大学)
Journal refRodríguez-Barroso, et. al. (2026). From Privacy to Trust in the Agentic Era: A Taxonomy of Challenges in Trustworthy Federated Learning Through the Lens of Trust Report 2.0. Information Fusion, 104236
机构
*
State Key Laboratory of AI Safety(人工智能安全国家重点实验室)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)