arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1721 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 越狱攻击 1721 篇

2106.07541 2021-06-17 math.OC cs.SY eess.SY 50%

Resilient Control of Platooning Networked Robotic Systems via Dynamic Watermarking

Matthew Porter, Arnav Joshi, Sidhartha Dey, Qirui Wu, Pedro Hespanhol, Anil Aswani, Matthew Johnson-Roberson, Ram Vasudevan

专题命中 越狱攻击 :safety(abstract)

Comments 19 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.06633 2021-02-15 eess.SY cs.SY 50%

Discrete-Time Consensus Networks: Scalability, Grounding and Countermeasures

Yamin Yan, Sonja Stüdli, Maria M. Seron, Richard H. Middleton

专题命中 越狱攻击 :safety(abstract)

Comments 12 pages,12 figures. arXiv admin note: substantial text overlap with arXiv:2002.11938

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.00615 2021-02-08 eess.SP cs.SY eess.SY 50%

Modeling Method for the Coupling Relations of Microgrid Cyber-Physical Systems Driven by Hybrid Spatiotemporal Events

Xiaoyong Bo, Xiaoyu Chen, Huashun Li, Yunchang Dong, Zhaoyang Qu, Lei Wang, Yang Li

专题命中 越狱攻击 :safety(abstract)

Comments Accepted by IEEE Access

Journal ref IEEE Access 9 (2021) 19619-19631

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.10106 2020-08-25 cs.CV 50%

Developing and Defeating Adversarial Examples

Ian McDiarmid-Sterling, Allan Moser

专题命中 越狱攻击 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.11938 2020-02-28 eess.SY cs.SY 50%

Analysis of Attack via Grounding and Countermeasures in Discrete-Time Consensus Networks

Yamin Yan, Sonja Stuedli, Maria M. Seron, Richard H. Middleton

专题命中 越狱攻击 :safety(abstract)

Comments 21st IFAC World Congress, 2020, to appear

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.06826 2019-08-21 cs.CR cs.CV eess.SP stat.ML 50%

Adversarial Sensor Attack on LiDAR-based Perception in Autonomous Driving

Yulong Cao, Chaowei Xiao, Benjamin Cyr, Yimeng Zhou, Won Park, Sara Rampazzi, Qi Alfred Chen, Kevin Fu, Z. Morley Mao

专题命中 越狱攻击 :safety(abstract)

Comments Accepted at the ACM Conference on Computer and Communications Security (CCS), 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.11026 2019-05-31 cs.CV cs.CR 50%

Fooling Detection Alone is Not Enough: First Adversarial Attack against Multiple Object Tracking

Yunhan Jia, Yantao Lu, Junjie Shen, Qi Alfred Chen, Zhenyu Zhong, Tao Wei

专题命中 越狱攻击 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.09969 2019-04-23 cs.CR 50%

Detecting ADS-B Spoofing Attacks using Deep Neural Networks

Xuhang Ying, Joanna Mazer, Giuseppe Bernieri, Mauro Conti, Linda Bushnell, Radha Poovendran

专题命中 越狱攻击 :safety(abstract)

Comments Accepted to IEEE CNS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.08970 2019-03-28 cs.CV 50%

Evaluation of Momentum Diverse Input Iterative Fast Gradient Sign Method (M-DI2-FGSM) Based Attack Method on MCS 2018 Adversarial Attacks on Black Box Face Recognition System

Md Ashraful Alam Milton

专题命中 越狱攻击 :safety(abstract)

Comments The Code is available for download in the following github link: https://github.com/miltonbd/mcs_2018_adversarial_attack . arXiv admin note: text overlap with arXiv:1803.06978 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.03639 2019-02-15 cs.CR 50%

Crossfire Attack Detection using Deep Learning in Software Defined ITS Networks

Akash Raj Narayanadoss, Tram Truong-Huu, Purnima Murali Mohan, Mohan Gurusamy

专题命中 越狱攻击 :safety(abstract)

Comments This paper has been accepted for publication in the proceeding of IEEE VTC2019-Spring

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.05447 2018-10-15 cs.CR 50%

How to Pick Your Friends - A Game Theoretic Approach to P2P Overlay Construction

Saar Tochner, Aviv Zohar

专题命中 越狱攻击 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏