arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 7971 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 7971 篇

2607.23523 2026-07-28 cs.AR 新提交 71%

CircuitWeave: Topology-Behavior Alignment for Executable Multimodal RTL Generation

CircuitWeave:用于可执行多模态RTL生成的拓扑-行为对齐

Jiahao Feng, Haiyan Qin, Zhiwei Xie, Wang Kang

专题命中 其他安全 :alignment(title)

AI总结 研究如何从自然语言规范生成RTL,提出CircuitWeave合同介导多模态框架,从原理图和文本提取合同并融合,通过联合目标监督相关过程,在生成可执行RTL上取得较好效果,部分指标高于无原理图的情况。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.19081 2026-06-18 q-bio.NC cs.HC 新提交 71%

Retrieval-Based Brain Decoding by Alignment, not Complexity

基于对齐而非复杂性的检索式脑解码

Matteo Ciferri, Matteo Ferrante, Nicola Toschi

专题命中 其他安全 :alignment(title)

AI总结 本文通过跨多数据集实验证明,线性对比解码器在脑解码中优于岭回归和标准非线性方法,表明解码增益更多来自训练目标而非架构复杂性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18441 2026-06-18 cs.CV 新提交 71%

Reasoning as Intersection: Consensus-Frame Alignment for Visual Focus in Video-MLLMs

推理即交集:视频多模态大语言模型中视觉焦点的一致性帧对齐

Chengwen Liu, Zhe Huang, Jisheng Dang, Hong Peng, Qi Tian, Tat-Seng Chua

机构 * School of Information Science and Engineering, Lanzhou University(兰州大学信息科学与工程学院) Beijing University of Posts and Telecommunications(北京邮电大学) Cloud and AI BU, Huawei(华为云与AI业务部) School of Computing, National University of Singapore(新加坡国立大学计算机学院)

专题命中 其他安全 :alignment(title)

AI总结 提出无时间标注的过程级奖励框架CF-GRPO,通过视频内在线索构建一致性帧先验,并利用一致性帧奖励优化模型帧使用与先验的对齐,提升视频推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.13190 2026-06-12 cs.RO cs.HC 新提交 71%

Multi-Modal Multi-Agent Robotic Cognitive Alignment enabled by Non-Invasive Consumer Brain Computer Interfaces: A Proof of Concept Exploration

基于非侵入式消费级脑机接口的多模态多智能体机器人认知对齐:概念验证探索

Nataliya Kosmyna, Liz Jenkins, Anoop K. Sinha

机构 * GOOGLE(谷歌) Paradigms of Intelligence(智能范式) Cambridge, MA, United States(马萨诸塞州剑桥市,美国) Mountain View, CA, United States(加利福尼亚州山景城,美国)

专题命中 其他安全 :alignment(title)

AI总结 提出一种框架,利用消费级脑机接口监测脑电信号,在高认知负荷时延迟智能体通信,实现认知对齐的多智能体交互,初步验证了实时信号处理、大语言模型与机器人结合的可行性。

Comments 19 pages, 9 figures, for associated video, see https://youtu.be/0Tav-G87XGs

详情

展开后加载摘要…

URL PDF HTML 收藏
1611.01824 2026-06-04 eess.SY cs.SY 71%

Robust Distance-Based Formation Control of Multiple Rigid Bodies with Orientation Alignment

多刚体鲁棒距离基形成控制

Alexandros Nikou, Christos K. Verginis, Dimos V. Dimarogonas

专题命中 其他安全 :alignment(title)

AI总结 本文研究了在3D空间中,针对第二类非线性多智能体系统,设计了去中心化无模型控制协议,实现距离和方向基的形成控制,并通过仿真验证了控制器性能。

Comments IFAC Word Congress 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.01970 2026-06-02 cs.RO cs.MA cs.SY eess.SY 71%

Market-Based Replanning for Safety-Critical UAV Swarms in Search and Rescue Missions

基于市场重规划的搜救任务中安全关键无人机群

Luiz Giacomossi, Andrea Haglund, Claire Namatovu, Emily Zainali, Esaias Målqvist, Yonatan M. Beyene, Ivan Tomasic, Baran Çürüklü, Håkan Forsberg

机构 * KTH Royal Institute of Technology(皇家理工学院) Swedish Defence Research Agency(瑞典国防研究机构) KTH Royal Institute technological Institute(皇家理工学院)

专题命中 其他安全 :safety(title)

AI总结 提出一种分布式协调架构IRDS,通过反向拍卖市场机制和几何共识协议,在无人机故障下自主重分配任务,在25%退化下保持93%任务成功率。

Comments 6 pages, 4 figures, accepted at MIPRO 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01562 2026-05-22 quant-ph 71%

Kernel Alignment for Quantum Support Vector Machines Using Genetic Algorithms

使用遗传算法的量子支持向量机核的核对齐

Floyd M. Creevey, Jamie A. Heredge, Martin E. Sevior, Lloyd C. L. Hollenberg

专题命中 其他安全 :alignment(title)

AI总结 本文提出了一种基于遗传算法的量子支持向量机核对齐方法,通过评估监督和无监督核损失函数对编码电路优化的影响,提高了分类准确率,并在金融、医疗和材料科学等领域展示了改进的机器学习性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00429 2026-04-14 eess.SY cs.SY 71%

Distributed Safety-Critical Control of Multi-Agent Systems with Time-Varying Communication Topologies

多智能体系统中具有时间变化通信拓扑的分布式安全控制

Shiyu Cheng, Luyao Niu, Bhaskar Ramasubramanian, Andrew Clark, Radha Poovendran

专题命中 其他安全 :safety(title)

AI总结 本文提出一种基于分布式优化的控制框架,通过截断函数和辅助不匹配变量处理时间变化通信拓扑下的多智能体协同与避障问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25460 2026-03-27 cs.SD 71%

CLAR: CIF-Localized Alignment for Retrieval-Augmented Speech LLM-Based Contextual ASR

CLAR:基于检索增强的语音大语言模型的上下文语音识别中的CIF局部对齐

Shangkun Huang, Huan Shen, Wei Zou, Yunzhang Chen

机构 * BRVoice Team, Bairong, Inc., China(BRVoice团队,百融云创,中国)

专题命中 其他安全 :alignment(title)

AI总结 CLAR通过CIF学习单调性token级对齐,提升语音识别中专有名词和长尾词的识别效果,减少表示稀释和注意力漂移,提高检索准确率。

Comments Submitted to Interspeech 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01477 2026-03-03 cs.RO 71%

SFCo-Nav: Efficient Zero-Shot Visual Language Navigation via Collaboration of Slow LLM and Fast Attributed Graph Alignment

SFCo-Nav: 通过慢LLM与快速属性图对齐的协作实现高效的零样本视觉语言导航

Chaoran Xiong, Litao Wei, Xinhao Hu, Kehui Ma, Ziyi Xia, Zixin Jiang, Zhen Sun, Ling Pei

机构 * Shanghai Key Laboratory of Navigation and Location Based Services, Shanghai Jiao Tong University(上海导航与位置基于服务重点实验室,上海交通大学)

专题命中 其他安全 :alignment(title)

AI总结 SFCo-Nav通过慢LLM与快速属性图对齐的协作,实现了高效的零样本视觉语言导航,显著提升效率并降低计算成本。

Comments Accepted by 2026 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23068 2026-02-27 cs.SD 71%

TADA: A Generative Framework for Speech Modeling via Text-Acoustic Dual Alignment

TADA: 一种通过文本-语音双模对齐生成语音模型的框架

Trung Dang, Sharath Rao, Ananya Gupta, Christopher Gagne, Panagiotis Tzirakis, Alice Baird, Jakub Piotr Cłapa, Peter Chin, Alan Cowen

机构 * Hume AI Dartmouth College(达特茅斯学院)

专题命中 其他安全 :alignment(title)

AI总结 TADA通过文本-语音双模对齐生成框架,实现语音与文本的一对一同步建模,提升语音生成的保真度和效率,减少幻觉并降低推理成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.10593 2026-02-12 cs.CV 71%

Fast Person Detection Using YOLOX With AI Accelerator For Train Station Safety

利用YOLOX与AI加速器的快速人员检测用于车站安全

Mas Nurul Achmadiah, Novendra Setyawan, Achmad Arif Bryantono, Chi-Chia Sun, Wen-Kai Kuo

机构 * Department of Electro-Optics, National Formosa University, Taiwan(国立Formosa大学电子光学系) Department of Electrical Engineering, National Taipei University, Taiwan(国立台北大学电子工程系) Department of Electrical Engineering, University of Muhammadiyah Malang, Indonesia(穆罕默迪亚大学Malang分校电子工程系) Department of Electronics Engineering, State Polytechnic of Malang, Indonesia(Malang州立理工学院电子工程系) Smart Manufacturing and Intelligent Machinery Research Center, National Formosa University, Taiwan(国立Formosa大学智能制造与智能机械研究中心) Department of Electronics Engineering, National Formosa University, Taiwan(国立Formosa大学电子工程系)

专题命中 其他安全 :safety(title)

AI总结 本文提出利用YOLOX与Hailo-8 AI加速器提高车站乘客检测的准确性和效率

Comments 6 pages, 8 figures, 2 tables. Presented at 2024 International Electronics Symposium (IES). IEEE DOI: 10.1109/IES63037.2024.10665874

Journal ref 2024 International Electronics Symposium (IES), pp. 504-509, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.08559 2026-02-10 cs.IR 71%

QARM V2: Quantitative Alignment Multi-Modal Recommendation for Reasoning User Sequence Modeling

QARM V2:量化对齐多模态推荐用于推理用户序列建模

Tian Xia, Jiaqi Zhang, Yueyang Liu, Hongjian Dou, Tingya Yin, Jiangxia Cao, Xulei Liang, Tianlu Xie, Lihao Liu, Xiang Chen, Shen Wang, Changxin Lao, Haixiang Gan, Jinkai Yu, Keting Cen, Lu Hao, Xu Zhang, Qiqiang Zhong, Zhongbo Sun, Yiyu Wang, Shuang Yang, Mingxin Wen, Xiangyu Wu, Shaoguo Liu, Tingting Gao, Zhaojie Liu, Han Li, Kun Gai

专题命中 其他安全 :alignment(title)

AI总结 QARM V2通过量化对齐多模态推荐方法,解决推荐系统中用户序列建模的语义理解与业务需求不匹配问题。

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20418 2026-02-04 eess.IV cs.CV 71%

Diff4MMLiTS: Advanced Multimodal Liver Tumor Segmentation via Diffusion-Based Image Synthesis and Alignment

Diff4MMLiTS: 通过基于扩散的图像合成与对齐的先进多模态肝肿瘤分割

Shiyun Chen, Li Lin, Pujin Cheng, ZhiCheng Jin, JianJian Chen, HaiDong Zhu, Kenneth K. Y. Wong, Xiaoying Tang

机构 * Department of Electronic and Electrical Engineering, Southern University of Science and Technology, Shenzhen, China(电子与电气工程系,南方科技大学,深圳,中国) Department of Electrical and Electronic Engineering, The University of Hong Kong, Hong Kong SAR, China(电气与电子工程系,香港大学,香港特别行政区,中国) Department of Radiology, Zhongda Hospital, Medical School, Southeast University, Nanjing, China(放射科,中大医院,医学院,东南大学,南京,中国) Jiaxing Research Institute, Southern University of Science and Technology, Jiaxing, China(嘉兴研究所,南方科技大学,嘉兴,中国)

专题命中 其他安全 :alignment(title)

AI总结 Diff4MMLiTS通过基于扩散的图像合成与对齐技术,实现肝肿瘤的多模态分割,无需严格对齐的多模态数据,提升了分割性能。

Comments International Workshop on Machine Learning in Medical Imaging, 668-678

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01268 2026-02-03 cs.CV cs.RO 71%

OASIS-DC: Generalizable Depth Completion via Output-level Alignment of Sparse-Integrated Monocular Pseudo Depth

OASIS-DC: 通过稀疏集成单目伪深度的输出级对齐实现通用深度补全

Jaehyeon Cho, Jhonghyun An

机构 * Vehicle Intelligence Perception Lab (VIPLAB), Gachon University, Seongnam-si, Republic of Korea(高银大学)

专题命中 其他安全 :alignment(title)

AI总结 OASIS-DC通过输出级对齐稀疏集成单目伪深度,实现无需大量标注样本的通用深度补全。

Comments Accepted to ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17627 2026-01-27 cs.SE 71%

Code Change Characteristics and Description Alignment: A Comparative Study of Agentic versus Human Pull Requests

代码变更特征与描述对齐:代理与人类拉取请求的比较研究

Dung Pham, Taher A. Ghaleb

专题命中 其他安全 :alignment(title)

AI总结 研究比较了代理与人类生成的拉取请求在代码变更特征和描述质量上的差异,发现代理在提交级消息质量上表现更好,但在PR级总结上不如人类,揭示了代理在微观精确性与宏观沟通之间的差距。

Comments Accepted at the 23rd International Conference on Mining Software Repositories (MSR '26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.18117 2025-12-23 cs.IR 71%

Factorized Transport Alignment for Multimodal and Multiview E-commerce Representation Learning

因子化运输对多模态和多视角电商表示学习

Xiwen Chen, Yen-Chieh Lien, Susan Liu, María Castaños, Abolfazl Razi, Xiaoting Zhao, Congzhe Su

专题命中 其他安全 :alignment(title)

AI总结 本文提出因子化运输框架,通过统一多模态和多视角学习,提升电商场景下的检索性能。

Comments Accepted by WSDM'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04555 2025-12-08 cs.RO cs.CV 71%

Evo-1: Lightweight Vision-Language-Action Model with Preserved Semantic Alignment

Evo-1:轻量级视觉-语言-动作模型,保持语义对齐

Tao Lin, Yilei Zhong, Yuxin Du, Jingjing Zhang, Jiting Liu, Yinxinyu Chen, Encheng Gu, Ziyan Liu, Hongyi Cai, Yanwen Zou, Lixing Zou, Zhaoye Zhou, Gen Li, Bo Zhao

机构 * School of AI, Shanghai Jiao Tong University(上海交通大学人工智能学院) EvoMind Tech(EvoMind科技) IAAR-Shanghai(IAAR-上海) SII Carnegie Mellon University(卡内基梅隆大学) University of Cambridge(剑桥大学) Nanyang Technological University(南洋理工大学)

专题命中 其他安全 :alignment(title)

AI总结 Evo-1是一种轻量级的视觉-语言-动作模型,通过减少计算并保持语义对齐,实现了高效的部署和强大的性能。

Comments Github: https://github.com/MINT-SJTU/Evo-1

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05718 2025-11-10 cs.HC 71%

Do Vision-Language Models See Visualizations Like Humans? Alignment in Chart Categorization

Péter Ferenc Gyarmati, Manfred Klaffenböck, Laura Koesten, Torsten Möller

专题命中 其他安全 :alignment(title)

Comments 2 pages, 2 figures. Accepted submission to the poster track of IEEE VIS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18835 2025-11-05 cs.SE 71%

AUCAD: Automated Construction of Alignment Dataset from Log-Related Issues for Enhancing LLM-based Log Generation

Hao Zhang, Dongjun Yu, Lei Zhang, Guoping Rong, Yongda Yu, Haifeng Shen, He Zhang, Dong Shao, Hongyu Kuang

专题命中 其他安全 :alignment(title)

Comments In the 16th International Conference on Internetware 2025. 13 pages

Journal ref Proceedings of the 16th International Conference on Internetware (2025) 413-425

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12205 2025-10-15 eess.SY cs.SY 71%

Sleepy Chauffeur Detection and Alert Techniques for Road Safety

Himel Ghosh, Sayak Chatterjee, Antik Ganguly, Shreetama Karmakar, Koushik Sarkar

专题命中 其他安全 :safety(title)

Comments 8 pages, 5 figures, International Journal on Recent Innovation in Microelectronics and Microcontrollers Applications Vol. 1, Issue 1 - 2018

Journal ref International Journal on Recent Innovation in Microelectronics and Microcontrollers Applications Vol. 1, Issue 1 - 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15548 2025-09-30 cs.SD 71%

Versatile Symbolic Music-for-Music Modeling via Function Alignment

Junyan Jiang, Daniel Chin, Liwei Lin, Xuanjie Liu, Gus Xia

专题命中 其他安全 :alignment(title)

Journal ref The 26th conference of the International Society for Music Information Retrieval (ISMIR 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19472 2025-09-25 eess.SY cs.SY 71%

Automatic and Scalable Safety Verification using Interval Reachability with Subspace Sampling

Brendan Gould, Akash Harapanahalli, Samuel Coogan

专题命中 其他安全 :safety(title)

Comments 6 pages, 3 figures. Updated to correct a small error in the dynamics presented in equation (12)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15940 2025-09-22 cs.DC 71%

Efficient Pre-Training of LLMs via Topology-Aware Communication Alignment on More Than 9600 GPUs

Guoliang He, Youhe Jiang, Wencong Xiao, Kaihua Jiang, Shuguang Wang, Jun Wang, Zixian Du, Zhuo Jiang, Xinlei Zhang, Binhang Yuan, Eiko Yoneki

专题命中 其他安全 :alignment(title)

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08820 2025-09-11 cs.RO 71%

RoboChemist: Long-Horizon and Safety-Compliant Robotic Chemical Experimentation

Zongzheng Zhang, Chenghao Yue, Haobo Xu, Minwen Liao, Xianglin Qi, Huan-ang Gao, Ziwei Wang, Hao Zhao

专题命中 其他安全 :safety(title)

Comments Accepted to CoRL 2025, Project Page: https://zzongzheng0918.github.io/RoboChemist.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03687 2025-08-05 cs.IR 71%

Fine-grained Alignment of Large Language Models for General Medication Recommendation without Overprescription

Zihao Zhao, Chenxiao Fan, Junlong Liu, Zheng Wang, Xiangnan He, Chongming Gao, Juan Li, Fuli Feng

专题命中 其他安全 :alignment(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09139 2025-07-15 cs.CV 71%

PoseLLM: Enhancing Language-Guided Human Pose Estimation with MLP Alignment

Dewen Zhang, Tahir Hussain, Wangpeng An, Hayaru Shouno

机构 * Department of Informatics, Graduate School of Informatics and Engineering, The University of Electro-Communications(信息学院,信息工程研究生院,东京电通大学)

专题命中 其他安全 :alignment(title)

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14437 2025-06-18 cs.IR 71%

Similarity = Value? Consultation Value Assessment and Alignment for Personalized Search

Weicong Qin, Yi Xu, Weijie Yu, Teng Shi, Chenglei Shen, Ming He, Jianping Fan, Xiao Zhang, Jun Xu

专题命中 其他安全 :alignment(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.12159 2025-05-27 cs.MA 71%

An Identity Based Agent Model for Value Alignment

Karthik Sama, Janvi Chhabra, Arpitha Srivatsha Malavalli, Jayati Deshmukh, Srinath Srinivasa

专题命中 其他安全 :alignment(title)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05217 2025-04-08 cs.IR 71%

LLM-Alignment Live-Streaming Recommendation

Yueyang Liu, Jiangxia Cao, Shen Wang, Shuang Wen, Xiang Chen, Xiangyu Wu, Shuang Yang, Zhaojie Liu, Kun Gai, Guorui Zhou

专题命中 其他安全 :alignment(title)

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏