机构
*
University of Southern California(南加州大学)
;
Iowa State University(爱荷华州立大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
UT Austin(德克萨斯大学奥斯汀分校)
;
Independent Researcher(独立研究员)
;
University of Notre Dame(圣母大学)
机构
*
Tsinghua University(清华大学)
;
Beijing Normal University(北京师范大学)
;
South China University of Technology(华南理工大学)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Shenzhen ShenNong Information Technology Co., Ltd.(深圳神农信息技术有限公司)
Adaptive and Explicit safe: Triggering Latent Safety Awareness in Large Reasoning Models
自适应且显式安全:触发大型推理模型中的潜在安全意识
Ke Miao, Jiaxin Li, Hongliang Chen, Yuke Hu, Zhan Qin
机构
*
The State Key Laboratory of Blockchain and Data Security, Zhejiang University(浙江大学区块链与数据安全全国重点实验室)
;
Hangzhou HighTech Zone (Binjiang) Blockchain and Data Security Research Institute, China(杭州高新区(滨江)区块链与数据安全研究院)
;
Li Auto Inc.(理想汽车)
;
Tsinghua University(清华大学)
;
King Abdullah University Of Science And Technology(阿卜杜拉国王科技大学)
NeuroSymbolic AI for Legal AI-TRISM: Trustworthy, Reliable, Interpretable, Safe Models
面向法律AI-TRISM的神经符号AI:可信、可靠、可解释、安全模型
Deepa Tilwani, Yash Saxena, Ankur Padia, Srinivasan Parthasarathy, Manas Gaur
机构
*
Department of Computer Science, AI Institute, University of South Carolina(南卡罗来纳大学计算机科学系,人工智能研究所)
;
Department of Computer Science and Electrical Engineering, University of Maryland, Baltimore County(马里兰大学巴尔的摩县分校计算机科学与电气工程系)
;
Department of Computer Science and Engineering, The Ohio State University(俄亥俄州立大学计算机科学与工程系)
机构
*
Qwen Large Model Application Team, Alibaba(阿里巴巴通义千问大模型应用团队)
;
Renmin University of China(中国人民大学)
;
Peking University(北京大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Zurich(苏黎世大学)
;
The Chinese University of Hong Kong(香港中文大学)
Distributed Safe Consensus Under Asymmetric Input and Time-Varying Output Constraints
非对称输入与时变输出约束下的分布式安全一致性
Abhinav Sinha, Shashi Ranjan Kumar
机构
*
Guidance, Autonomy, Learning, and Control for Intelligent Systems (GALACxIS) Lab, Department of Aerospace Engineering and Engineering Mechanics, University of Cincinnati(智能系统引导、自主、学习与控制实验室,航空航天工程与工程力学系,辛辛那提大学)
;
Intelligent Systems and Control (ISaC) Lab, Department of Aerospace Engineering, Indian Institute of Technology Bombay(智能系统与控制实验室,航空航天工程系,印度班加罗尔理工学院)
DoubtProbe: Black-Box Jailbreak Defense via Structural Verification and Semantic Auditing
DoubtProbe:通过结构验证与语义审计的黑盒越狱防御
Xuanyu Yin, Yilin Jiang, Jun Zhou, Kai Chen, Zhengfu Cao, Xiaolei Dong
机构
*
East China Normal University(东华大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
机构
*
East China Normal University(东华大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
University of Southampton(南安普顿大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Software Engineering Institute, East China Normal University(东华大学软件工程研究院)
GAS-Leak-LLM: Genetic Algorithm-Based Suffix Optimization for Black-Box LLM Jailbreaking
GAS-Leak-LLM:基于遗传算法的后缀优化实现黑盒LLM越狱
Aman Anifer, Vignesh Kumar Kembu, Vishnu M, Antonino Nocera, Vinod P., Amal Murali PK, Akshay S Rajan
机构
*
Department of Electrical, Computer and Biomedical Engineering(电气、计算机与生物医学工程系)
;
University of Pavia(帕维亚大学)
;
Department of Computer Applications(计算机应用系)
;
Cochin University of Science and Technology(科钦科学技术大学)