Machine Learning in Micromobility: A Systematic Review of Datasets, Techniques, and Applications
专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG
Comments 14 pages, 3 tables, and 4 figures, submitted to IEEE Transactions on Intelligent Vehicles
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG
Comments 14 pages, 3 tables, and 4 figures, submitted to IEEE Transactions on Intelligent Vehicles
机构 * Department of Computer Science, Anglia Ruskin University(计算机科学系,安格利亚 Ruskin 大学)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.LG
Comments 13 pages. Mechanistic analysis of backdoored LLMs (Qwen2.5-3B). Code: https://github.com/mshahoyi/sa_attn_analysis. Base model: unsloth/Qwen2.5-3B-Instruct-unsloth-bnb-4bit. Finetuned models: https://huggingface.co/collections/mshahoyi/simple-sleeper-agents-68a1df3a7aaff310aa0e5336
机构 * Shanghai Jiao Tong University(上海交通大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments Work in progress
机构 * NSF Center for Quantum Network(NSF量子网络中心) ; University of California, Los Angeles(加州大学洛杉矶分校)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments ICML 2025 Workshop on MAS
机构 * Stanford University(斯坦福大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
Journal ref Transactions on Machine Learning Research (TMLR) 2835-8856 (2025)
机构 * LARG, Research Center for Social Computing and Interactive Robotics, Harbin Institute of Technology(大型语言模型研究组,社会计算与交互机器人研究中心,哈尔滨工业大学) ; School of Computer Science and Engineering, Central South University(计算机科学与工程学院,中南大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments Preprint
机构 * OpenAI
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI
机构 * Department of Computing, The Hong Kong Polytechnic University(计算机系,香港理工大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted by IEEE TKDE. Codes and data are available at https://github.com/Quhaoh233/TokenRec
机构 * ISIR, Sorbonne Université(ISIR,索邦大学)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI
Comments ICCV 2025. The first three authors contributed equally. Project page and code: https://pegah- kh.github.io/projects/lmm-finetuning-analysis-and-steering/
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
机构 * Institute for People-Centred AI and Centre for Translation Studies, School of Computer Science and Electronic Engineering, University of Surrey(以人为本的人工智能研究所和翻译研究中心,计算机科学与电子工程学院,萨里大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted to COLM 2025 Conference
机构 * School of Computer Science, Faculty of Engineering, University of Sydney(悉尼大学计算机科学学院、工程学院) ; Computational Health Informatics Program, Boston Children’s Hospital(波士顿儿童医院计算健康信息学项目) ; Harvard-MIT Center for Regulatory Science and Department of Pediatrics, Harvard Medical School(哈佛-麻省理工监管科学中心和哈佛医学院儿科部门) ; Sydney School of Public Health, Faculty of Medicine and Health, University of Sydney(悉尼大学公共卫生学院、医学与健康学院)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI
Comments 12 pages, 4 figures. Updated to include Table 2, Supplementary Table 1, and an additional baseline random forest model
机构 * Walmart Global Tech(沃尔玛全球科技)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
机构 * Charité – Universitätsmedizin Berlin, Department of Psychiatry and Psychotherapy, Berlin, Germany(柏林查理医院医学大学精神病与心理治疗系) ; Hertie Institute for AI in Brain Health, University of Tübingen, Germany(图宾根大学健康人工智能研究所)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
Comments Accepted to ICML 2025
机构 * Purdue University(普渡大学) ; Johns Hopkins University(约翰霍普金斯大学)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI
Comments COLM 2025. The first two authors contributed equally to this work
机构 * Independent Researcher in AI and Statistics(人工智能与统计学独立研究者) ; Shahrood University of Technology(沙霍罗德大学) ; University of Pittsburgh(匹兹堡大学) ; Duquesne University(杜克森大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments 6pages
机构 * Center for Human-Compatible AI, University of California, Berkeley(人类兼容人工智能中心,加州大学伯克利分校) ; Foundations of Cooperative AI Lab, Carnegie Mellon University(协作人工智能实验室,卡内基梅隆大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship
专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG
Comments arXiv admin note: substantial text overlap with arXiv:2505.04480
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
Comments Code is available at https://github.com/furiosa-ai/uncage
机构 * Department of Computer Science, University of Maryland, Baltimore County(计算机科学系,马里兰大学巴尔的摩县分校) ; Department of Dermatology, Johns Hopkins University School of Medicine(皮肤科系,约翰霍普金斯大学医学院)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
Comments Accepted to IJCAI 2025 Workshops. arXiv admin note: substantial text overlap with arXiv:2506.10328
机构 * Aalto University(阿alto大学) ; Finnish Center for Artificial Intelligence (FCAI)(芬兰人工智能中心) ; System 2 AI
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG
Comments Preprint
机构 * Sapienza University of Rome(罗马大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG
Comments Accepted as Full Research Papers at CIKM 2025
机构 * Universidad Nacional de Colombia(哥伦比亚国立大学)
专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.LG
Comments Based on master's thesis in Systems and Computer Engineering, Universidad Nacional de Colombia (2025)
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; ShanghaiTech University(上海科技大学) ; University of Hong Kong(香港大学) ; Fudan University(复旦大学)
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG
Comments Accepted by the 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025), Main Conference
Journal ref Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2025
机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) ; University of Waterloo(滑铁卢大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted in UIST 2025
专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG
Comments To appear in CoRL 2025
专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG