arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 7968 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 7968 篇

1909.05006 2024-07-23 q-bio.PE cond-mat.stat-mech cs.LG q-bio.BM stat.ML 74%

Boltzmann machine learning and regularization methods for inferring evolutionary fields and couplings from a multiple sequence alignment

Sanzo Miyazawa

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments In arXiv:1909.05006v3 the values of selective temperature for protein PF00153, $T_s$ in Table 5 and in the section 2.8, and folding free energy for PF00595, and in the v4 the method for soft-thresholding were corrected; shown in red. The v2 was published in the IEEE/ACM Transactions on Computational Biology and Bioinformatics. The program is available from https://gitlab.com/sanzo.miyazawa/BM/

Journal ref IEEE/ACM Transactions on Computational Biology and Bioinformatics, 2022 Jan-Feb;19(1):328-342

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.10245 2024-07-16 cs.CL cs.IR 74%

GenSco: Can Question Decomposition based Passage Alignment improve Question Answering?

Barah Fazili, Koustava Goswami, Natwar Modani, Inderjeet Nair

专题命中 其他安全 :alignment(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.19570 2024-07-15 cs.CR cs.AI 74%

Synthetic Cancer -- Augmenting Worms with LLMs

Benjamin Zimmerman, David Zollikofer

专题命中 其他安全 :safety(abstract,comments);AI safety(abstract,comments);分类 cs.AI

Comments Won first place at the Swiss AI Safety Prize. Some technical details omitted, contact authors for more information

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.18073 2024-05-29 cs.AI 74%

Towards Dialogues for Joint Human-AI Reasoning and Value Alignment

Elfia Bezou-Vrakatseli, Oana Cocarascu, Sanjay Modgil

专题命中 其他安全 :alignment(title);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.13702 2024-04-23 astro-ph.CO astro-ph.GA cs.LG 74%

Learning Galaxy Intrinsic Alignment Correlations

Sneh Pandya, Yuanyuan Yang, Nicholas Van Alfen, Jonathan Blazek, Robin Walters

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments 15 pages, 6 figures, 1 table. Accepted at the Data-centric Machine Learning Research (DMLR) Workshop at ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.01575 2024-04-19 cs.LG cs.HC stat.ML 74%

StackGenVis: Alignment of Data, Algorithms, and Models for Stacking Ensemble Learning Using Performance Metrics

Angelos Chatzimparmpas, Rafael M. Martins, Kostiantyn Kucher, Andreas Kerren

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments This manuscript is accepted for publication in a special issue of IEEE Transactions on Visualization and Computer Graphics Journal (IEEE TVCG)

Journal ref IEEE TVCG 2021, 27(2), 1547-1557

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05696 2024-03-01 cs.CL 74%

A Preliminary Study of the Intrinsic Relationship between Complexity and Alignment

Yingxiu Zhao, Bowen Yu, Binyuan Hui, Haiyang Yu, Fei Huang, Yongbin Li, Nevin L. Zhang

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments LREC-Coling 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13577 2024-02-22 cs.CL 74%

BBA: Bi-Modal Behavioral Alignment for Reasoning with Large Vision-Language Models

Xueliang Zhao, Xinting Huang, Tingchen Fu, Qintong Li, Shansan Gong, Lemao Liu, Wei Bi, Lingpeng Kong

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13181 2024-02-21 cs.RO cs.LG 74%

DINOBot: Robot Manipulation via Retrieval and Alignment with Vision Foundation Models

Norman Di Palo, Edward Johns

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments To appear at 2024 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06960 2023-12-13 cs.CV cs.LG 74%

Remote Sensing Vision-Language Foundation Models without Annotations via Ground Remote Alignment

Utkarsh Mall, Cheng Perng Phoo, Meilin Kelsey Liu, Carl Vondrick, Bharath Hariharan, Kavita Bala

专题命中 其他安全 :alignment(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13061 2023-11-23 cs.CL 74%

Attribution and Alignment: Effects of Local Context Repetition on Utterance Production and Comprehension in Dialogue

Aron Molnar, Jaap Jumelet, Mario Giulianelli, Arabella Sinclair

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments CoNLL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03792 2023-11-08 cs.CL 74%

Character-Level Bangla Text-to-IPA Transcription Using Transformer Architecture with Sequence Alignment

Jakir Hasan, Shrestha Datta, Ameya Debnath

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Achieved top position with a word error rate of 0.10582 in the public ranking of DataVerse Challenge - ITVerse 2023 (link: https://www.kaggle.com/competitions/dataverse_2023/). All codes can be found on the respective competition webpage

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05454 2023-11-07 cs.CL 74%

Flesch or Fumble? Evaluating Readability Standard Alignment of Instruction-Tuned Language Models

Joseph Marvin Imperial, Harish Tayyar Madabushi

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Final camera-ready for EMNLP GEM Workshop 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.06443 2023-08-15 cs.LG eess.AS 74%

Neural Latent Aligner: Cross-trial Alignment for Learning Representations of Complex, Naturalistic Neural Data

Cheol Jun Cho, Edward F. Chang, Gopala K. Anumanchipalli

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments Accepted at ICML 2023

Journal ref Proceedings of the 40th International Conference on Machine Learning (2023), PMLR 202:5661-5676

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11952 2023-05-23 cs.CL 74%

Self-QA: Unsupervised Knowledge Guided Language Model Alignment

Xuanyu Zhang, Qing Yang

专题命中 其他安全 :alignment(title);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.08535 2022-11-10 cs.CL 74%

The challenges of temporal alignment on Twitter during crises

Aniket Pramanick, Tilman Beck, Kevin Stowe, Iryna Gurevych

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Accepted to Findings of EMNLP, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.04482 2022-03-31 cs.CV cs.CL 74%

FLAVA: A Foundational Language And Vision Alignment Model

Amanpreet Singh, Ronghang Hu, Vedanuj Goswami, Guillaume Couairon, Wojciech Galuba, Marcus Rohrbach, Douwe Kiela

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.14133 2022-01-13 cs.LG 74%

Self-Labeling of Fully Mediating Representations by Graph Alignment

Martijn Oldenhof, Adam Arany, Yves Moreau, Jaak Simm

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments Code available: https://github.com/biolearning-stadius/chemgrapher-self-rich-labeling

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.05862 2021-11-22 cs.LG stat.ML 74%

Constrained Non-Affine Alignment of Embeddings

Yuwei Wang, Yan Zheng, Yanqing Peng, Chin-Chia Michael Yeh, Zhongfang Zhuang, Das Mahashweta, Bendre Mangesh, Feifei Li, Wei Zhang, Jeff M. Phillips

专题命中 其他安全 :alignment(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.04316 2021-11-10 cs.CV cs.LG cs.NE 74%

COVID-19 Face Mask Recognition with Advanced Face Cut Algorithm for Human Safety Measures

Arkaprabha Basu, Md Firoj Ali

专题命中 其他安全 :safety(title);分类 cs.LG

Comments 5 pages, 7 figures

Journal ref 2021 12th International Conference on Computing Communication and Networking Technologies (ICCCNT)

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12028 2021-09-27 cs.CL 74%

Investigating Post-pretraining Representation Alignment for Cross-Lingual Question Answering

Fahim Faisal, Antonios Anastasopoulos

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Accepted at MRQA Workshop 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.12547 2021-04-13 cs.CL 74%

Multilingual BERT Post-Pretraining Alignment

Lin Pan, Chung-Wei Hang, Haode Qi, Abhishek Shah, Saloni Potdar, Mo Yu

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Accepted at NAACL2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.00195 2020-12-02 cs.LG q-bio.BM 74%

Profile Prediction: An Alignment-Based Pre-Training Task for Protein Sequence Models

Pascal Sturmfels, Jesse Vig, Ali Madani, Nazneen Fatema Rajani

专题命中 其他安全 :alignment(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.03590 2020-11-30 cs.RO cs.LG cs.SY eess.SY 74%

Reactive motion planning with probabilistic safety guarantees

Yuxiao Chen, Ugo Rosolia, Chuchu Fan, Aaron D. Ames, Richard Murray

专题命中 其他安全 :safety(title);分类 cs.LG

Comments In the Conference on Robotic Learning 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.01056 2020-01-07 cs.LG stat.AP stat.ML 74%

Root Cause Detection Among Anomalous Time Series Using Temporal State Alignment

Sayan Chakraborty, Smit Shah, Kiumars Soltani, Anna Swigart

专题命中 其他安全 :alignment(title);分类 cs.LG

Comments 6 pages, 7 figures, 2019 18th IEEE International Conference on Machine Learning and Applications (ICMLA)

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.05727 2019-11-14 cs.CY cs.IR eess.IV 74%

Artificial Intelligence Strategies for National Security and Safety Standards

Erik Blasch, James Sung, Tao Nguyen, Chandra P. Daniel, Alisa P. Mason

专题命中 其他安全 :safety(title);分类 cs.CY

Comments Presented at AAAI FSS-19: Artificial Intelligence in Government and Public Sector, Arlington, Virginia, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.03562 2019-05-22 cs.LG stat.ML 74%

Real time Traffic Flow Parameters Prediction with Basic Safety Messages at Low Penetration of Connected Vehicles

Mizanur Rahman, Mashrur Chowdhury, Jerome McClendon

专题命中 其他安全 :safety(title);分类 cs.LG

Comments 16 pages, 15 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.08129 2018-11-21 cs.IR cs.CL 74%

Alignment Analysis of Sequential Segmentation of Lexicons to Improve Automatic Cognate Detection

Pranav A

专题命中 其他安全 :alignment(title);分类 cs.CL

Comments Published at ACL-SRW 2018

Journal ref Proceedings of ACL 2018, Student Research Workshop. 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.03479 2018-03-12 cs.AI 74%

Highly Automated Learning for Improved Active Safety of Vulnerable Road Users

Maarten Bieshaar, Günther Reitberger, Viktor Kreß, Stefan Zernetsch, Konrad Doll, Erich Fuchs, Bernhard Sick

专题命中 其他安全 :safety(title);分类 cs.AI

Comments 4 pages, 1 figure

Journal ref published in ACM Chapters Computer Science in Cars Symposium (CSCS-17). Munich, Germany. 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
1608.07398 2016-08-29 cs.LO cs.AI cs.SE 74%

Proceedings First Workshop on Causal Reasoning for Embedded and safety-critical Systems Technologies

Gregor Gössler, Oleg Sokolsky

专题命中 其他安全 :safety(title);分类 cs.AI

Journal ref EPTCS 224, 2016

详情

展开后加载摘要…

URL PDF HTML 收藏