Time-Series Forecasting in Safety-Critical Environments: An Open-Source Package for EU-AI-Act-Compliant Development / Zeitreihenprognose in sicherheitskritischen Umgebungen: Ein Open-Source-Paket für die KI-VO-konforme Entwicklung
CommentsAccepted manuscript: Workshop on Emerging Behaviors in Embodied AI for Achieving Robust Autonomy as part of European Conference on Computer Vision (ECCV) 2026
Enhance the Safety in Reinforcement Learning by ADRC Lagrangian Methods
通过ADRC拉格朗日方法增强强化学习的安全性
Mingxu Zhang, Huicheng Zhang, Jiaming Ji, Yaodong Yang, Ying Sun
机构
*
AI Thrust, The Hong Kong University of Science and Technology (Guangzhou)(人工智能方向,香港科技大学(广州))
;
School of Artificial Intelligence, Peking University, Beijing, China(人工智能学院,北京大学,北京,中国)
;
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
机构
*
School of Cyber Science and Technology, Beihang University, Beijing, China(北京航空航天大学网络安全学院)
;
Institute of Artificial Intelligence, Beihang University, Beijing, China(北京航空航天大学人工智能研究院)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学)
;
AI Security Lab, Beijing, China(360人工智能安全实验室)
Comments4 pages, 3 figures. Accepted at the 1st IJCAI Workshop on Safe Physical AI (SPAI 2026), held in conjunction with IJCAI-ECAI 2026, Bremen, Germany
Selective Safety Steering via Value-Filtered Decoding
基于价值过滤解码的选择性安全引导
Bat-Sheva Einbinder, Hen Davidov, Yee Whye Teh, Yarin Gal, Yaniv Romano
机构
*
Department of Electrical and Computer Engineering, Technion IIT(技术学院电气与计算机工程系)
;
Department of Statistics, University of Oxford(牛津大学统计系)
;
OATML, Department of Computer Science, University of Oxford(牛津大学计算机科学系)
;
Department of Computer Science, Technion IIT(技术学院计算机科学系)