arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1836 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1836 篇

2007.12123 2021-10-19 cs.RO math.LO math.OC 50%

Receding Horizon Control Based Online Motion Planning with Partially Infeasible LTL Specifications

Mingyu Cai, Hao Peng, Zhijun Li, Hongbo Gao, Zhen Kan

专题命中 AI治理与伦理 :safety(abstract)

Journal ref IEEE Control Systems Letters, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.00971 2021-02-02 q-bio.BM 50%

Methodology-centered review of molecular modeling, simulation, and prediction of SARS-CoV-2

Kaifu Gao, Rui Wang, Jiahui Chen, Limei Cheng, Jaclyn Frishcosy, Yuta Huzumi, Yuchi Qiu, Tom Schluckbier, Guo-Wei Wei

专题命中 AI治理与伦理 :safety(abstract)

Comments 99 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.16116 2020-08-19 astro-ph.CO astro-ph.GA 50%

Patterns of galaxy spin directions in SDSS and Pan-STARRS show parity violation and multipoles

Lior Shamir

专题命中 AI治理与伦理 :alignment(abstract)

Comments ApSS, accepted. arXiv admin note: substantial text overlap with arXiv:1912.05429

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.05429 2019-12-18 astro-ph.GA astro-ph.CO 50%

Large-scale patterns of galaxy spin rotation show cosmological-scale parity violation and multipoles

Lior Shamir

专题命中 AI治理与伦理 :alignment(abstract)

Comments To be submitted. Comments welcome

详情

展开后加载摘要…

URL PDF HTML 收藏
1705.02321 2017-05-08 cs.GT 50%

Fairness Incentives for Myopic Agents

Sampath Kannan, Michael Kearns, Jamie Morgenstern, Mallesh Pai, Aaron Roth, Rakesh Vohra, Z. Steven Wu

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
0912.3984 2010-01-14 cs.MA 50%

Multi-Agent Model using Secure Multi-Party Computing in e-Governance

Durgesh Kumar Mishra, Samiksha Shukla

专题命中 AI治理与伦理 :safety(abstract)

Journal ref Journal of Computing, Volume 1, Issue 1, pp 195-199, December 2009

详情

展开后加载摘要…

URL PDF HTML 收藏