arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

AAAI Conference on Artificial Intelligence · 会议 · Artificial Intelligence

2026-07-30 至 2026-07-30 共收录 7
2607.26339 2026-07-30 cs.LG cs.CR cs.IR 新提交

RAGuard: A Layered Defense Framework for Retrieval-Augmented Generation Systems Against Data Poisoning

RAGuard:一种针对检索增强生成系统数据中毒的分层防御框架

Pushkal Kumar, Tucker Nielson, Tanish Kolhe, Shubham Zala, Vincent Li

AI总结 该研究针对RAG系统的语料库中毒攻击,提出分层防御框架RAGuard,含对抗微调检索器与ZKIP两层,可将攻击成功率降至0,且保持检索性能,相关资源已开源。

Comments Accepted to NeurIPS ResponsibleFM 2025, AAAI FrontierIR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26119 2026-07-30 cs.AI cs.CL 新提交

Probing the Origins of Reasoning Performance: Representational Quality for Mathematical Problem-Solving in RL vs. SFT Fine-Tuned Models

探究推理性能的起源:强化学习(RL)与监督微调(SFT)模型在数学问题求解中的表征质量

Antyabha Rahman, Akshaj Gurugubelli, Omar Ankit, Kevin Zhu, Aishwarya Balwani

AI总结 本研究探究RL与SFT微调模型数学推理性能差异的机制,发现RL模型的表征更具线性可分性、深层重要性更高,其token分配变异性或反映在线策略推理的分布。

Comments Second Workshop on XAI4Science, AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11178 2026-07-30 cs.AI cs.CL cs.MM cs.SI 版本更新

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

TANDEM: 面向多模态仇恨言论的时间感知神经检测

Girish A. Koushik, Helen Treharne, Diptesh Kanojia

机构 * Nature-Inspired Computing & Engineering, University of Surrey(Surrey大学自然启发计算与工程系) Surrey Centre for Cyber Security, University of Surrey(Surrey大学网络安全中心)

AI总结 提出TANDEM统一框架,通过串联强化学习策略联合优化视觉-语言和音频-语言模型,将音频-视觉仇恨检测转化为结构化推理问题,在HateMM上目标识别F1达0.73(提升30%),并保持精确时间定位。

Comments Accepted to AAAI-ICWSM 2027

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04480 2026-07-30 cs.AI cs.CE cs.SY eess.SY math.OC

Prescriptive Artificial Intelligence: A Formal Paradigm for Auditing Human Decisions Under Uncertainty

规范性人工智能:一种用于在不确定性下审计人类决策的正式范式

Pedro Passos

机构 * Institute of Computing, Universidade Federal Fluminense (UFF)(联邦弗鲁米嫩塞大学计算研究所)

AI总结 本文提出规范性人工智能作为高风险环境下的人类-人工智能决策协作范式,通过四个公理证明了其与预测系统的核心区别,并展示了在决策模仿中系统性偏差的限制。

Comments Preprint; suitable for AI, decision sciences, and prescriptive analytics. Short versions published in Wharton Sports Analytics Journal Fall 2025 (AI Feature Spotlight) and accepted to AAAI Bridge on LM Reasoning 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14265 2026-07-30 cs.CV

Bayesian Neural Networks for One-to-Many Mapping in Image Enhancement

基于贝叶斯神经网络的图像增强中的一对多映射

Guoxi Huang, Qirui Yang, Ruirui Lin, Zipeng Qi, David Bull, Nantheera Anantrasirichai

AI总结 本文提出贝叶斯增强模型,利用贝叶斯神经网络和确定性神经网络解决图像增强中的一对多映射问题,通过捕捉数据不确定性生成多样化输出。

Journal ref Proceedings of the AAAI conference on artificial intelligence (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07210 2026-07-30 cs.CV cs.CR cs.LG 版本更新

Breaking the Stealth-Potency Trade-off in Clean-Image Backdoors with Generative Trigger Optimization

利用生成式触发器优化打破干净图像后门的隐蔽性-有效性权衡

Binyan Xu, Fan Yang, Di Tang, Xilin Dai, Kehuan Zhang

AI总结 该研究提出生成式干净图像后门框架GCB,利用条件InfoGAN优化触发器,仅需极小投毒示例即可实现攻击,CA下降不到1%,适配多类数据集、架构与任务,且能抵御多数后门防御。

Comments 19 pages, 22 figures, 15 tables. To appear in AAAI '26 (Oral). This paper extends the AAAI-2026 version by including the Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14766 2026-07-30 cs.CV cs.CL 版本更新

ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM

ASCD:用于减少多模态大语言模型(MLLM)幻觉的注意力可导向对比解码

Yujun Wang, Aniri, Jinhe Bi, Soeren Pirk, Yunpu Ma

AI总结 该研究针对MLLM的幻觉问题,提出ASCD方法,通过正负引导调整解码时的注意力分数,在多基准上显著减少幻觉并提升VQA准确率,且无需额外训练。

Comments Accepted at AAAI 2026

Journal ref Proceedings of the AAAI Conference on Artificial Intelligence, 40(12): 10306-10314, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏