Towards Secure Logging: Characterizing and Benchmarking Logging Code Security Issues with LLMs
面向安全日志:利用LLM分析和基准测试日志代码的安全问题
He Yang Yuan, Xin Wang, Kundi Yao, An Ran Chen, Zishuo Ding, Zhenhao Li
机构
*
York University(约克大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
University of Waterloo(滑铁卢大学)
;
University of Alberta(阿尔伯塔大学)
RExBench: Can coding agents autonomously implement AI research extensions?
RExBench:代码代理能否自主实现AI研究扩展?
Nicholas Edwards, Yukyung Lee, Yujun Audrey Mao, Yulu Qin, Sebastian Schuster, Najoung Kim
机构
*
Faculty of Computer Science, University of Vienna(维也纳大学计算机科学系)
;
UniVie Doctoral School Computer Science, University of Vienna(维也纳大学UniVie计算机科学博士学院)
;
Boston University(波士顿大学)
专题命中
评测与基准
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
Comments20 pages. Preprint; arXiv long version of a paper accepted at AIware 2026. Adds Appendices A (cross-language) and B (Python isolation) not present in the ACM camera-ready
Assessing the Robustness of Climate Foundation Models under No-Analog Distribution Shifts
在无类比分布偏移下评估气候基础模型的鲁棒性
Maria Conchita Agana Navarro, Geng Li, Theo Wolf, Maria Perez-Ortiz
机构
*
Centre for Artificial Intelligence(人工智能中心)
;
Department of Computer Science(计算机科学系)
;
University College London(伦敦大学学院)
;
Division of Emerging Interdisciplinary Areas(新兴交叉学科领域)
;
Hong Kong University of Science and Technology(香港科技大学)
;
University of Oxford(牛津大学)