arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 7978 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 7978 篇

2504.14287 2025-04-22 cs.CL cs.CY 62%

Probing the Subtle Ideological Manipulation of Large Language Models

Demetris Paschalides, George Pallis, Marios D. Dikaiakos

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05805 2025-04-22 cs.LG cs.AI cs.MA 62%

Multi-agent Auto-Bidding with Latent Graph Diffusion Models

Dom Huh, Prasant Mohapatra

机构 * Department of Computer Science University of California, Davis(计算机科学系加州大学戴维斯分校)

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.06353 2025-04-22 cs.LG cs.AI cs.CV 62%

Deep Active Learning in the Open World

Tian Xie, Jifan Zhang, Haoyue Bai, Robert Nowak

机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校)

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19887 2025-04-17 cs.CY cs.AI 62%

AI threats to national security can be countered through an incident regime

Alejandro Ortega

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.19950 2025-04-17 cs.LG cs.AI 62%

Multimodal Lego: Model Merging and Fine-Tuning Across Topologies and Modalities in Biomedicine

Konstantin Hemker, Nikola Simidjievski, Mateja Jamnik

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14779 2025-04-17 cs.CY cs.AI cs.HC 62%

The Use of Generative Artificial Intelligence for Upper Secondary Mathematics Education Through the Lens of Technology Acceptance

Mika Setälä, Ville Heilala, Pieta Sikström, Tommi Kärkkäinen

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.CY

Comments Published in the Proceedings of the 40th ACM/SIGAPP Symposium on Applied Computing (SAC'25), March 31--April 4, 2025, Catania, Italy

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09967 2025-04-15 cs.CV cs.AI cs.LG 62%

Enhancing Multi-task Learning Capability of Medical Generalist Foundation Model via Image-centric Multi-annotation Data

Xun Zhu, Fanbin Mo, Zheng Zhang, Jiaxi Wang, Yiming Shi, Ming Wu, Chuang Zhang, Miao Li, Ji Wu

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09680 2025-04-15 cs.LG cs.AI math.OC 62%

SPOT: Spatio-Temporal Pattern Mining and Optimization for Load Consolidation in Freight Transportation Networks

Sikai Cheng, Amira Hijazi, Jeren Konak, Alan Erera, Pascal Van Hentenryck

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05424 2025-04-15 cs.CL cs.AI 62%

SAMGPT: Text-free Graph Foundation Model for Multi-domain Pre-training and Cross-domain Adaptation

Xingtong Yu, Zechuan Gong, Chang Zhou, Yuan Fang, Hui Zhang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments Accepted by WWW2025 Main Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08849 2025-04-15 cs.CY cs.AI cs.CE 62%

Exploring Cognitive Attributes in Financial Decision-Making

Mallika Mainali, Rosina O. Weber

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.CY

Comments 7 pages, 2 figures. Presented in SIAM International Conference on Data Mining (SDM25) METACOG-25: 2nd Workshop on Metacognitive Prediction of AI Behavior

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08415 2025-04-14 cs.LG cs.AI 62%

Constrained Machine Learning Through Hyperspherical Representation

Gaetano Signorelli, Michele Lombardi

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19346 2025-04-11 cs.CV cs.CL cs.LG 62%

CLIP meets DINO for Tuning Zero-Shot Classifier using Unlabeled Image Collections

Mohamed Fazli Imam, Rufael Fedaku Marew, Jameel Hassan, Mustansar Fiaz, Alham Fikri Aji, Hisham Cholakkal

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03994 2025-04-09 cs.LG cs.AI cs.MA cs.SY eess.SY 62%

Improving Mixed-Criticality Scheduling with Reinforcement Learning

Muhammad El-Mahdy, Nourhan Sakr, Rodrigo Carrasco

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

Comments This work was submitted to the 32nd International Conference on Real-Time Networks and Systems (RTNS) on June 8, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02885 2025-04-09 cs.CV cs.AI cs.LG 62%

Expertized Caption Auto-Enhancement for Video-Text Retrieval

Baoyao Yang, Junxiang Chen, Wanyun Li, Wenbin Yao, Yang Zhou

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07300 2025-04-09 cs.LG cs.CL 62%

CALF: Aligning LLMs for Time Series Forecasting via Cross-modal Fine-Tuning

Peiyuan Liu, Hang Guo, Tao Dai, Naiqi Li, Jigang Bao, Xudong Ren, Yong Jiang, Shu-Tao Xia

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05019 2025-04-08 cs.LG cs.CL 62%

Mixture-of-Personas Language Models for Population Simulation

Ngoc Bui, Hieu Trung Nguyen, Shantanu Kumar, Julian Theodore, Weikang Qiu, Viet Anh Nguyen, Rex Ying

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04238 2025-04-08 cs.CL cs.AI 62%

Sensitivity Meets Sparsity: The Impact of Extremely Sparse Parameter Patterns on Theory-of-Mind of Large Language Models

Yuheng Wu, Wentao Guo, Zirui Liu, Heng Ji, Zhaozhuo Xu, Denghui Zhang

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02607 2025-04-04 cs.LG cs.AI cs.SY eess.SY 62%

Learning Geometrically-Informed Lyapunov Functions with Deep Diffeomorphic RBF Networks

Samuel Tesfazgi, Leonhard Sprandl, Sandra Hirche

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01039 2025-04-03 cs.CY cs.AI 62%

One Person, One Bot

Liat Lavi

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.CY

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10114 2025-04-02 cs.LG cs.CL cs.CV 62%

Mixture of Experts Made Personalized: Federated Prompt Learning for Vision-Language Models

Jun Luo, Chen Chen, Shandong Wu

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.LG

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21888 2025-03-31 cs.CL cs.AI 62%

RedditESS: A Mental Health Social Support Interaction Dataset -- Understanding Effective Social Support to Refine AI-Driven Support Tools

Zeyad Alghamdi, Tharindu Kumarage, Garima Agrawal, Mansooreh Karami, Ibrahim Almuteb, Huan Liu

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07900 2025-03-27 cs.CL cs.AI 62%

High-Dimension Human Value Representation in Large Language Models

Samuel Cahyawijaya, Delong Chen, Yejin Bang, Leila Khalatbari, Bryan Wilie, Ziwei Ji, Etsuko Ishii, Pascale Fung

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19123 2025-03-26 cs.CL cs.AI 62%

Overcoming Vocabulary Mismatch: Vocabulary-agnostic Teacher Guided Language Modeling

Haebin Shin, Lei Ji, Xiao Liu, Yeyun Gong

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15405 2025-03-26 cs.CL cs.AI 62%

Semantic Layered Embedding Diffusion in Large Language Models for Multi-Contextual Consistency

Irin Kabakum, Thomas Montgomery, Daniel Ravenwood, Genevieve Harrington

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

Comments arXiv admin note: This paper has been withdrawn by arXiv due to disputed and unverifiable authorship

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09387 2025-03-25 cs.CV cs.AI cs.LG 62%

RankCLIP: Ranking-Consistent Language-Image Pretraining

Yiming Zhang, Zhuokai Zhao, Zhaorun Chen, Zhili Feng, Zenghui Ding, Yining Sun

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments Code and model checkpoints are available at https://github.com/Jam1ezhang/RankCLIP

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17803 2025-03-25 cs.LG cs.AI cs.MA stat.ME 62%

A Roadmap Towards Improving Multi-Agent Reinforcement Learning With Causal Discovery And Inference

Giovanni Briglia, Stefano Mariani, Franco Zambonelli

专题命中 其他安全 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01238 2025-03-24 cs.AI cs.CL eess.SP 62%

Large Language Models are Zero-Shot Recognizers for Activities of Daily Living

Gabriele Civitarese, Michele Fiori, Priyankar Choudhary, Claudio Bettini

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI

Comments Paper accepted for publication in the ACM Transactions on Intelligent Systems and Technology (TIST) journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11898 2025-03-18 cs.CL cs.AI 62%

LLMs for Translation: Historical, Low-Resourced Languages and Contemporary AI Models

Merve Tekgurler

专题命中 其他安全 :safety(abstract);分类 cs.CL、cs.AI

Comments Accepted to LaTeCH-CLfL 2025, held in conjunction with NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.07207 2025-03-18 cs.CV cs.AI cs.CL 62%

Valley: Video Assistant with Large Language model Enhanced abilitY

Ruipu Luo, Ziwang Zhao, Min Yang, Zheming Yang, Minghui Qiu, Tao Wang, Zhongyu Wei, Yanhao Wang, Cen Chen

专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.12961 2025-03-18 cs.LG cs.AI 62%

Position: Tensor Networks are a Valuable Asset for Green AI

Eva Memmel, Clara Menzen, Jetze Schuurmans, Frederiek Wesel, Kim Batselier

专题命中 其他安全 :alignment(abstract);分类 cs.AI、cs.LG

Comments This paper has been accepted for presentation at the International Conference on Machine Learning (ICML) 2024 and will appear in the conference proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏