arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 3248 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 安全训练 3248 篇

2409.09845 2025-08-05 cs.RO 78%

Friction-Aware Safety Locomotion for Wheeled-legged Robots using Vision Language Models and Reinforcement Learning

Bo Peng, Donghoon Baek, Qijie Wang, Joao Ramos

机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Department of Electrical and Computer Engineering(电气与计算机工程系) Department of Mechanical Science Engineering(机械科学与工程系) School of Software(软件学院)

专题命中 安全训练 :safety(title,abstract)

Comments Accepted to Humanoids 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23226 2025-08-01 cs.CV 78%

Toward Safe, Trustworthy and Realistic Augmented Reality User Experience

Yanming Xiu

机构 * Department of Electrical and Computer Engineering, Duke University(电子工程系,杜克大学)

专题命中 安全训练 :trustworthy(title);safety(abstract)

Comments 2 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20685 2025-07-29 eess.SY cs.SY 78%

What's Really Different with AI? -- A Behavior-based Perspective on System Safety for Automated Driving Systems

Marcus Nolte, Nayel Fabian Salem, Olaf Franke, Jan Heckmann, Christoph Höhmann, Georg Stettinger, Markus Maurer

专题命中 安全训练 :safety(title,abstract)

Comments 8 pages, 1 figure, 1 table, to be published in 2025 IEEE International Automated Vehicle Validation Conference (IAVVC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04980 2025-07-16 cs.RO cs.SY eess.SY 78%

LVLM-MPC Collaboration for Autonomous Driving: A Safety-Aware and Task-Scalable Control Architecture

Kazuki Atsuta, Kohei Honda, Hiroyuki Okuda, Tatsuya Suzuki

机构 * Department of Mechanical Systems Engineering(机械系统工程系) Nagoya University(名古屋大学)

专题命中 安全训练 :safety(title,abstract)

Comments 8 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22894 2025-07-01 cs.RO 78%

Safe Reinforcement Learning with a Predictive Safety Filter for Motion Planning and Control: A Drifting Vehicle Example

Bei Zhou, Baha Zarrouki, Mattia Piccinini, Cheng Hu, Lei Xie, Johannes Betz

机构 * State Key Laboratory of Industrial Control Technology, Zhejiang University(浙江大学工业控制技术国家重点实验室) Professorship of Autonomous Vehicle Systems, Technical University of Munich(慕尼黑工业大学自主车辆系统教授职位)

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02620 2025-06-04 eess.SY cs.RO cs.SY 78%

Back to Base: Towards Hands-Off Learning via Safe Resets with Reach-Avoid Safety Filters

Azra Begzadić, Nikhil Uday Shinde, Sander Tonkens, Dylan Hirsch, Kaleb Ugalde, Michael C. Yip, Jorge Cortés, Sylvia Herbert

专题命中 安全训练 :safety(title,abstract)

Comments The first three authors contributed equally to the work

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19860 2025-05-27 cs.RO 78%

Causal Bayesian Networks for Data-driven Safety Analysis of Complex Systems

Roman Gansch, Lina Putze, Tjark Koopmann, Jan Reich, Christian Neurohr

机构 * Robert Bosch GmbH, Corporate Research(罗伯特·博世有限公司,企业研究) German Aerospace Center (DLR) e.V., Institute of Systems Engineering for Future Mobility(德国航空航天中心(DLR)协会,未来交通系统工程研究所) Fraunhofer Institute for Experimental Software Engineering (IESE)(弗劳恩霍夫实验软件工程研究所)

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.00973 2025-05-06 eess.SY cs.SY 78%

Defense Strategies for Autonomous Multi-agent Systems: Ensuring Safety and Resilience Under Exponentially Unbounded FDI Attacks

Yichao Wang, Mohamadamin Rajabinezhad, Dimitra Panagou, Shan Zuo

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04153 2025-04-21 cs.RO math.OC 78%

A Dynamic Safety Shield for Safe and Efficient Reinforcement Learning of Navigation Tasks

Murad Dawood, Ahmed Shokry, Maren Bennewitz

专题命中 安全训练 :safety(title,abstract)

Comments Accepted in L4DC2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06670 2025-04-10 cs.RO 78%

Dynamic Residual Safe Reinforcement Learning for Multi-Agent Safety-Critical Scenarios Decision-Making

Kaifeng Wang, Yinsong Chen, Qi Liu, Xueyuan Li, Xin Gao

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14656 2025-03-20 cs.RO math.OC 78%

Safety-Critical and Distributed Nonlinear Predictive Controllers for Teams of Quadrupedal Robots

Basit Muhammad Imran, Jeeseop Kim, Taizoon Chunawala, Alexander Leonessa, Kaveh Akbari Hamed

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06892 2025-03-11 cs.RO 78%

SafePlan: Leveraging Formal Logic and Chain-of-Thought Reasoning for Enhanced Safety in LLM-based Robotic Task Planning

Ike Obi, Vishnunandan L. N. Venkatesh, Weizheng Wang, Ruiqi Wang, Dayoon Suh, Temitope I. Amosa, Wonse Jo, Byung-Cheol Min

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13697 2025-01-24 eess.SY cs.SY stat.ML 78%

Safety in safe Bayesian optimization and its ramifications for control

Christian Fiedler, Johanna Menn, Sebastian Trimpe

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16809 2025-01-08 cs.CE 78%

Analytical assessment of workers' safety concerning direct and indirect ways of getting infected by dangerous pathogen

Krzysztof Domino, Arkadiusz Sochan, Jarosław Adam Miszczak

专题命中 安全训练 :safety(title,abstract)

Journal ref Journal of Computational Science Volume 85, February 2025, 102509

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.15437 2024-12-23 eess.SY cs.SY 78%

Safety-Critical Control of Discontinuous Systems with Nonsmooth Safe Sets

Mohammed Alyaseen, Nikolay Atanasov, Jorge Cortes

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10031 2024-11-18 eess.SY cs.SY 78%

Enforcing Cooperative Safety for Reinforcement Learning-based Mixed-Autonomy Platoon Control

Jingyuan Zhou, Longhao Yan, Jinhao Liang, Kaidi Yang

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03570 2024-11-04 cs.RO 78%

Embodied AI with Two Arms: Zero-shot Learning, Safety and Modularity

Jake Varley, Sumeet Singh, Deepali Jain, Krzysztof Choromanski, Andy Zeng, Somnath Basu Roy Chowdhury, Avinava Dubey, Vikas Sindhwani

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.03885 2024-10-08 cs.RO cs.SY eess.SY math.OC 78%

Collaborative Safety-Critical Formation Control with Obstacle Avoidance

Brooks A. Butler, Chi Ho Leung, Philip E. Paré

专题命中 安全训练 :safety(title,abstract)

Comments This work is under review for publication in Automatica. arXiv admin note: text overlap with arXiv:2311.11156

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15343 2024-09-25 cs.IR 78%

Advertiser Content Understanding via LLMs for Google Ads Safety

Joseph Wallace, Tushar Dogra, Wei Qiao, Yuan Wang

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11057 2024-09-25 cs.RO cs.MA 78%

Safety Guaranteed Robust Multi-Agent Reinforcement Learning with Hierarchical Control for Connected and Automated Vehicles

Zhili Zhang, H M Sabbir Ahmad, Ehsan Sabouni, Yanchao Sun, Furong Huang, Wenchao Li, Fei Miao

专题命中 安全训练 :safety(title,abstract)

Comments 6 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14580 2024-09-24 cs.RO 78%

Updating Robot Safety Representations Online from Natural Language Feedback

Leonardo Santos, Zirui Li, Lasse Peters, Somil Bansal, Andrea Bajcsy

专题命中 安全训练 :safety(title,abstract)

Comments Submitted to ICRA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.11882 2024-04-19 eess.SY cs.RO cs.SY 78%

Hybrid Navigation Acceptability and Safety

Benoit Clement, Marie Dubromel, Paulo E. Santos, Karl Sammut, Michelle Oppert, Feras Dayoub

专题命中 安全训练 :safety(title);trustworthy(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.04665 2024-04-03 cs.RO 78%

Improving safety in mixed traffic: A learning-based model predictive control for autonomous and human-driven vehicle platooning

Jie Wang, Zhihao Jiang, Yash Vardhan Pant

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01096 2024-04-02 cs.SE cs.PL 78%

Enabling Memory Safety of C Programs using LLMs

Nausheen Mohammed, Akash Lal, Aseem Rastogi, Subhajit Roy, Rahul Sharma

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17830 2024-03-27 cs.CV 78%

Assessment of Multimodal Large Language Models in Alignment with Human Values

Zhelun Shi, Zhipin Wang, Hongxing Fan, Zaibin Zhang, Lijun Li, Yongting Zhang, Zhenfei Yin, Lu Sheng, Yu Qiao, Jing Shao

专题命中 安全训练 :alignment(title,abstract)

Comments arXiv admin note: text overlap with arXiv:2311.02692

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.17078 2024-03-27 cs.IR 78%

Unsupervised Large Language Model Alignment for Information Retrieval via Contrastive Feedback

Qian Dong, Yiding Liu, Qingyao Ai, Zhijing Wu, Haitao Li, Yiqun Liu, Shuaiqiang Wang, Dawei Yin, Shaoping Ma

专题命中 安全训练 :alignment(title,abstract)

Comments Accepted by SIGIR24

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03429 2024-03-26 cs.CR 78%

Behavioral Authentication for Security and Safety

Cheng Wang, Hao Tang, Hangyu Zhu, Junhan Zheng, Changjun Jiang

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01576 2024-02-05 cs.RO 78%

Training Adversarial yet Safe Agent to Characterize Safety Performance of Highly Automated Vehicles

Minghao Zhu, Anmol Sidhu, Keith A. Redmill

专题命中 安全训练 :safety(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.06245 2024-01-15 math.OC 78%

Distributed Optimal Output Consensus Control of Heterogeneous Multi-Agent Systems with Safety Constraints

Ji Ma, Shu Liang, Yiguang Hong

专题命中 安全训练 :safety(title,abstract)

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08689 2023-12-15 cs.RO 78%

Safety-Critical Coordination of Legged Robots via Layered Controllers and Forward Reachable Set based Control Barrier Functions

Jeeseop Kim, Jaemin Lee, Aaron D. Ames

专题命中 安全训练 :safety(title,abstract)

Comments 7 pages, 7 figures. arXiv admin note: substantial text overlap with arXiv:2303.13630

详情

展开后加载摘要…

URL PDF HTML 收藏