arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 7945 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 7945 篇

2210.04365 2022-11-11 cs.MA cs.AI cs.LG 81%

ELIGN: Expectation Alignment as a Multi-Agent Intrinsic Reward

Zixian Ma, Rose Wang, Li Fei-Fei, Michael Bernstein, Ranjay Krishna

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

Comments This paper will be published in Neurips 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.08540 2022-10-18 cs.CL cs.AI 81%

TransAlign: Fully Automatic and Effective Entity Alignment for Knowledge Graphs

Rui Zhang, Xiaoyan Zhao, Bayu Distiawan Trisedya, Min Yang, Hong Cheng, Jianzhong Qi

专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.13349 2022-06-28 cs.AI cs.CL 81%

Process Knowledge-Infused AI: Towards User-level Explainability, Interpretability, and Safety

Amit Sheth, Manas Gaur, Kaushik Roy, Revathy Venkataraman, Vedant Khandelwal

专题命中 其他安全 :safety(title,abstract);分类 cs.CL、cs.AI

Comments To paper in IEEE Internet Computing 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.08107 2022-06-17 cs.LG cs.AI 81%

Closed-Form Diffeomorphic Transformations for Time Series Alignment

Iñigo Martinez, Elisabeth Viles, Igor G. Olaizola

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

Comments 37 pages, 24 figures, 4 tables. Accepted at International Conference on Machine Learning ICML 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02841 2022-06-08 cs.CY cs.AI 81%

Researching Alignment Research: Unsupervised Analysis

Jan H. Kirchner, Logan Smith, Jacques Thibodeau, Kyle McDonell, Laria Reynolds

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.08507 2022-05-11 cs.AI cs.HC cs.LG cs.MA 81%

Towards an AI Coach to Infer Team Mental Model Alignment in Healthcare

Sangwon Seo, Lauren R. Kennedy-Metz, Marco A. Zenati, Julie A. Shah, Roger D. Dias, Vaibhav V. Unhelkar

专题命中 其他安全 :alignment(title);safety(abstract);分类 cs.AI、cs.LG

Comments Submitted to the 2021 IEEE Conference on Cognitive and Computational Aspects of Situation Management (CogSIMA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.00520 2022-05-06 cs.LG cs.AI 81%

The Role of Explainability in Assuring Safety of Machine Learning in Healthcare

Yan Jia, John McDermid, Tom Lawton, Ibrahim Habli

专题命中 其他安全 :safety(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.02863 2022-05-05 cs.LG cs.AI 81%

PocketNN: Integer-only Training and Inference of Neural Networks via Direct Feedback Alignment and Pocket Activations in Pure C++

Jaewoo Song, Fangzhen Lin

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

Comments Accepted in tinyML Research Symposium '22, March 2022, San Jose, CA (TinyML 2022). 7 pages, 4 figures, 2 tables. [v6] title is reverted to the original version

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.04823 2022-03-09 cs.LG cs.AI 81%

Taxonomy of Machine Learning Safety: A Survey and Primer

Sina Mohseni, Haotao Wang, Zhiding Yu, Chaowei Xiao, Zhangyang Wang, Jay Yadawa

专题命中 其他安全 :safety(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.03538 2022-03-08 cs.CL cs.AI 81%

AI-based Approach for Safety Signals Detection from Social Networks: Application to the Levothyrox Scandal in 2017 on Doctissimo Forum

Valentin Roche, Jean-Philippe Robert, Hanan Salam

专题命中 其他安全 :safety(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.00969 2022-01-05 cs.CV cs.CL cs.LG 81%

Interactive Attention AI to translate low light photos to captions for night scene understanding in women safety

Rajagopal A, Nirmala V, Arun Muthuraj Vedamanickam

专题命中 其他安全 :safety(title,abstract);分类 cs.CL、cs.LG

Comments In Springer Proceedings. International Conference On Big Data, Machine Learning and Applications 2021. http://bigdml.nits.ac.in/

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.03229 2021-10-22 cs.AI cs.LG 81%

Towards Sample Efficient Agents through Algorithmic Alignment

Mingxuan Li, Michael L. Littman

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.09980 2021-08-24 cs.CV cs.AI cs.LG 81%

TACo: Token-aware Cascade Contrastive Learning for Video-Text Alignment

Jianwei Yang, Yonatan Bisk, Jianfeng Gao

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

Comments Accepted by ICCV 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.03451 2021-07-26 cs.CL cs.AI 81%

Anticipating Safety Issues in E2E Conversational AI: Framework and Tooling

Emily Dinan, Gavin Abercrombie, A. Stevie Bergman, Shannon Spruit, Dirk Hovy, Y-Lan Boureau, Verena Rieser

专题命中 其他安全 :safety(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.04000 2021-07-09 cs.LG cs.AI cs.CV cs.RO 81%

Active Safety Envelopes using Light Curtains with Probabilistic Guarantees

Siddharth Ancha, Gaurav Pathak, Srinivasa G. Narasimhan, David Held

专题命中 其他安全 :safety(title,abstract);分类 cs.AI、cs.LG

Comments 18 pages, Published at Robotics: Science and Systems (RSS) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.09580 2021-04-21 cs.CL cs.AI cs.CV 81%

Improving Cross-Modal Alignment in Vision Language Navigation via Syntactic Information

Jialu Li, Hao Tan, Mohit Bansal

专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI

Comments NAACL 2021 (10 pages)

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.02625 2021-03-12 cs.SE cs.CY cs.LG 81%

Safety Case Templates for Autonomous Systems

Robin Bloomfield, Gareth Fletcher, Heidy Khlaaf, Luke Hinde, Philippa Ryan

专题命中 其他安全 :safety(title,abstract);分类 cs.CY、cs.LG

Comments 136 pages, 57 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.07306 2021-02-16 cs.CY cs.LG 81%

Vehicle to Vehicle (V2V) Communication Protocol: Components, Benefits, Challenges, Safety and Machine Learning Applications

Ramya Daddanala, Vekata Mannava, Lo'ai Tawlbeh, Mohammad Al-Ramahi

专题命中 其他安全 :safety(title,abstract);分类 cs.CY、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.11905 2020-09-28 cs.AI cs.LG cs.RO 81%

A New Approach for Tactical Decision Making in Lane Changing: Sample Efficient Deep Q Learning with a Safety Feedback Reward

M. Ugur Yavas, N. Kemal Ure, Tufan Kumbasar

专题命中 其他安全 :safety(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.01106 2020-02-11 cs.CY cs.CV cs.LG stat.ML 81%

Improving Traffic Safety Through Video Analysis in Jakarta, Indonesia

João Caldeira, Alex Fout, Aniket Kesari, Raesetje Sefala, Joseph Walsh, Katy Dupre, Muhammad Rizal Khaefi, Setiaji, George Hodge, Zakiya Aryana Pramestri, Muhammad Adib Imtiyazi

专题命中 其他安全 :safety(title,abstract);分类 cs.CY、cs.LG

Comments 6 pages; LaTeX; Presented at NeurIPS 2018 Workshop on Machine Learning for the Developing World; Presented at NeurIPS 2018 Workshop on AI for Social Good

Journal ref Proceedings of the 2019 Intelligent Systems Conference (IntelliSys) Volume 2, 642-649

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.02697 2019-04-05 cs.CY cs.AI 81%

A Systematic Literature Review about the impact of Artificial Intelligence on Autonomous Vehicle Safety

A. M. Nascimento, L. F. Vismari, C. B. S. T. Molina, P. S. Cugnasca, J. B. Camargo, J. R. de Almeida, R. Inam, E. Fersman, M. V. Marquezini, A. Y. Hata

专题命中 其他安全 :safety(title,abstract);分类 cs.AI、cs.CY

Comments 32 pages, 5 figures, 9 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.12185 2018-11-30 cs.CY cs.LG 81%

AI based Safety System for Employees of Manufacturing Industries in Developing Countries

Abhisek Das, Satanik Panda, Suman Datta, Soumitra Naskar, Pratep Misra, Tanushyam Chattopadhyay

专题命中 其他安全 :safety(title,abstract);分类 cs.CY、cs.LG

Comments Presented at NIPS 2018 Workshop on Machine Learning for the Developing World

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.09542 2018-06-26 cs.LG cs.CL stat.ML 81%

Mapping Unparalleled Clinical Professional and Consumer Languages with Embedding Alignment

Wei-Hung Weng, Peter Szolovits

专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.LG

Comments Accepted by 2018 KDD Workshop on Machine Learning for Medicine and Healthcare

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.06354 2018-02-07 cs.AI cs.HC cs.LG cs.RO 81%

Pragmatic-Pedagogic Value Alignment

Jaime F. Fisac, Monica A. Gates, Jessica B. Hamrick, Chang Liu, Dylan Hadfield-Menell, Malayandi Palaniappan, Dhruv Malik, S. Shankar Sastry, Thomas L. Griffiths, Anca D. Dragan

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

Comments Published at the International Symposium on Robotics Research (ISRR 2017)

Journal ref International Symposium on Robotics Research, 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
cs/0104007 2009-11-30 cs.LG cs.CL 81%

Bootstrapping Syntax and Recursion using Alignment-Based Learning

Menno van Zaanen

专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.LG

Comments 8 pages

Journal ref Proceedings of the Seventeenth International Conference on Machine Learning. pages 1063-1070

详情

展开后加载摘要…

URL PDF HTML 收藏
cs/0302015 2009-11-30 cs.AI cs.LG 81%

Unsupervised Learning in a Framework of Information Compression by Multiple Alignment, Unification and Search

J. G. Wolff

专题命中 其他安全 :alignment(title,abstract);分类 cs.AI、cs.LG

Comments 39 pages, 1 JPEG figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.17391 2026-04-21 cs.SE cs.AR cs.LG 80%

RISC-V Functional Safety for Autonomous Automotive Systems: An Analytical Framework and Research Roadmap for ML-Assisted Certification

RISC-V功能安全用于自动驾驶系统:面向ML辅助认证的分析框架与研究路线图

Nick Andreasyan, Mikhail Struve, Alexey Popov, Maksim Nikolaev, Vadim Vashkelis

机构 * Andes Technology(安德斯技术)

专题命中 其他安全 :safety(title,abstract);分类 cs.LG

AI总结 本文探讨RISC-V在自动驾驶功能安全中的角色,提出以认证经济性为核心的分析框架,通过ML方法优化认证流程,旨在构建符合ASIL-D标准的可认证RISC-V平台。

Comments 11 pages, 3 figures, 4 tables. Analytical perspective paper on automotive-grade RISC-V functional safety, certification economics, and ML-assisted certification for autonomous driving systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.17003 2025-04-08 cs.CR cs.AI 80%

Safety Layers in Aligned Large Language Models: The Key to LLM Security

Shen Li, Liuyi Yao, Lan Zhang, Yaliang Li

专题命中 其他安全 :safety(title,abstract);分类 cs.AI

Comments Accepted by ICLR 2025. The code is available at https://github.com/listen0425/Safety-Layers

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.07925 2025-01-23 cs.LG 80%

Phase of Flight Classification in Aviation Safety using LSTM, GRU, and BiLSTM: A Case Study with ASN Dataset

Aziida Nanyonga, Hassan Wasswa, Graham Wild

专题命中 其他安全 :safety(title,abstract);分类 cs.LG

Comments Aviation Safety, Deep learning algorithms, Flight phase, NLP, ASN, and Classification

Journal ref In 2023 International Conference on High Performance Big Data and Intelligent Systems (HDIS) (pp. 24-28). IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14890 2024-08-28 eess.AS cs.LG cs.SD 80%

Development of Large Annotated Music Datasets using HMM-based Forced Viterbi Alignment

S. Johanan Joysingh, P. Vijayalakshmi, T. Nagarajan

专题命中 其他安全 :alignment(title,abstract);分类 cs.LG

Comments submitted to TENCON 2019

Journal ref S. J. Joysingh, P. Vijayalakshmi and T. Nagarajan, "Development of Large Annotated Music Datasets using HMM based Forced Viterbi Alignment," TENCON 2019 - 2019 IEEE Region 10 Conference (TENCON), Kochi, India, 2019, pp. 1298-1302

详情

展开后加载摘要…

URL PDF HTML 收藏