TrustRoboReward: Preference-Ordered Isotonic Score Editing for Multi-Paradigm Robot Reward Models
TrustRoboReward:面向多范式机器人奖励模型的偏好有序保序分数编辑方法
Yidong Wang, Yan Zhan, Ziteng Feng, Zhenyu Cui, Ziyi Zhou, Renzhao Liang, Jiaxuan Zhu, Zilei Yang, Yiran Zhao, Zhongkuan Mao, Bo Jia, Hanchu Ni, Chenggang Xie, Biao Liu, Yi Zhang, Yong Dai, Xiaozhu Ju, Wei Ye, Shikun Zhang
机构
*
Peking University(北京大学)
;
Beijing Innovation Center of Humanoid Robotics(北京人形机器人创新中心)
;
University of Science and Technology of China(中国科学技术大学)
;
Southeast University(东南大学)
;
Southern University of Science and Technology(南方科技大学)
;
Beijing University of Aeronautics and Astronautics(北京航空航天大学)
;
Beijing Language and Culture University(北京语言大学)
;
Sichuan University(四川大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
Comments16 pages, 8 figures, 7 tables. To appear at CoNLL 2026
Journal refProceedings of the 30th Conference on Computational Natural Language Learning (CoNLL 2026), pp. 268-283, San Diego, California, USA, Association for Computational Linguistics, 2026
Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback
Wasserstein分布鲁棒遗憾优化用于人类反馈的强化学习
Yikai Wang, Shang Liu, Jose Blanchet
机构
*
Department of Statistics and Operations Research, University of North Carolina(统计与运筹学系,北卡罗来纳大学)
;
Imperial Business School, Imperial College London(帝国理工学院伦敦商学院)
;
Department of Management Science and Engineering, Stanford University(管理科学与工程系,斯坦福大学)
SHE: Trajectory-driven Safety Harness Evolution for LLM Agents
SHE:面向大语言模型智能体的轨迹驱动安全管控机制演化
Wanying Qu, Qinghua Mao, Yu Li, Jiyao Liu, Xin Zhang, Dadi Guo, Yanxu Zhu, Qingyu Liu, Leitao Yuan, Xi Lin, Shanfeng Zhu, Yanwei Fu, Jing Shao, Xia Hu, Dongrui Liu
机构
*
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Fudan University(复旦大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
Enhance the Safety in Reinforcement Learning by ADRC Lagrangian Methods
通过ADRC拉格朗日方法增强强化学习的安全性
Mingxu Zhang, Huicheng Zhang, Jiaming Ji, Yaodong Yang, Ying Sun
机构
*
AI Thrust, The Hong Kong University of Science and Technology (Guangzhou)(人工智能方向,香港科技大学(广州))
;
School of Artificial Intelligence, Peking University, Beijing, China(人工智能学院,北京大学,北京,中国)
;
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
Comments171 pages. Formalized in Lean 4 with Mathlib: 240 theorems in the elaborated environment, 141 audited headline results, cold-compiling from a clean checkout with zero custom axioms. Source, theorem-by-theorem contract, and reproducible axiom audit: https://github.com/selfreferencing/TSE_Formal. Companion to Agentic Capital