arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2509.18377 2026-05-29 cs.CL 81%

Interactive In-Meeting Speaker Correction with Human Feedback

交互式会议中基于人类反馈的说话人修正

Xinlu He, Yiwen Guan, Badrivishal Paurana, Pitipat Kongsomjit, Zilin Dai, Jacob Whitehill

机构 * Worcester Polytechnic Institute(沃斯特理工学院)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL

AI总结 提出一种LLM辅助的会议内说话人修正系统,通过用户简短反馈修正说话人归属错误,结合流式ASR、说话人日志、LLM摘要和在线注册机制,在AMI数据集上实现DER降低31.99%、说话人替换错误降低52.68%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27463 2026-05-28 stat.ME cs.AI stat.AP 81%

When prompt perturbations break your A/B test: A valid statistical test for generative surveying

当提示扰动破坏你的A/B测试:一种用于生成式调查的有效统计检验

Hayden Helm, Carey Priebe

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI

AI总结 针对生成式调查中LLM对提示设计敏感的问题,提出一种置换检验方法,在包含扰动结构的统计模型下保持有效性,并给出预算分配建议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26356 2026-05-27 cs.CL 81%

In-Context Optimization for Retrieval-Augmented Generation: A Gradient-Descent Perspective

检索增强生成的上下文优化:梯度下降视角

Mingchen Li, Jiatan Huang, Chuxu Zhang, Liang Zhao, Hong Yu

机构 * University of Massachusetts, Amherst(马萨诸塞大学阿默斯特分校) University of Connecticut(康涅狄格大学) Emory University(埃默里大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL

AI总结 本文从梯度下降视角研究检索增强生成(RAG)作为上下文优化过程,提出一种轻量级前向更新方法,在冻结LLM和检索器的情况下提升生成器对检索证据的利用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03545 2026-05-27 cs.AI 81%

Persona Generators: Generating Diverse Synthetic Personas for Arbitrary Contexts

人格生成器:为任意上下文生成多样化的合成人格

Davide Paglieri, Logan Cross, William A. Cunningham, Joel Z. Leibo, Alexander Sasha Vezhnevets

机构 * Google DeepMind(谷歌DeepMind)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出Persona Generators,通过迭代进化优化生成覆盖广泛意见和偏好的多样化合成人格,在六个多样性指标上显著优于现有基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23149 2026-05-27 cs.CL 81%

AlignEvoSkill: Towards Knowledge-Aware and Task-Aligned Agent Skill Evolution

AlignEvoSkill: 迈向知识感知与任务对齐的智能体技能进化

Dingzirui Wang, Xuanliang Zhang, Keyan Xu, Qingfu Zhu, Wanxiang Che, Yang Deng

机构 * Singapore Management University(新加坡国立大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL

AI总结 提出AlignEvoSkill框架,通过联合建模知识覆盖和任务对齐,从失败轨迹中识别知识标签、检索并适配候选技能,再基于知识覆盖和任务对齐分数筛选高质量技能,在3个基准和4个LLM骨干上相对提升34.7%,实现技能进化新SOTA且成本更低。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26040 2026-05-26 cs.AI 81%

L2IR: Revealing Latent Intent in Graph Fraud Detection

L2IR: 揭示图欺诈检测中的潜在意图

Jinsheng Guo, Zhenhao Weng, Yibo Liu, Yan Qiao, Meng Li

机构 * Hefei University of Technology(合肥工业大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出L2IR框架,利用大语言模型从用户行为和可疑连接中提取潜在意图,通过自适应自训练增强鲁棒性,在广泛伪装的数据集上提升图神经网络检测器的AUPRC最高达8.27%。

Comments 12 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12538 2026-05-26 cs.CR cs.AI cs.SE 81%

MCPXKIT: The Unified Toolkit for Analyzing Model Context Protocol Security

MCPXKIT:分析模型上下文协议安全性的统一工具包

Yongjian Guo, Puzhuo Liu, Wanlun Ma, Zehang Deng, Xiaogang Zhu, Peng Di, Xi Xiao, Sheng Wen

机构 * Shenzhen International Graduate School, Tsinghua University, Shenzhen, China(深圳国际研究生院,清华大学,深圳,中国) Ant Group, Hangzhou, China(蚂蚁集团,杭州,中国) Swinburne University of Technology, Melbourne, Australia(斯威本科技大学,墨尔本,澳大利亚) The University of Adelaide, Adelaide, Australia(阿德莱德大学,阿德莱德,澳大利亚) UNSW Sydney, Australia(悉尼大学,澳大利亚)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出MCPXKIT工具包,分类实现了31种攻击方法,通过定量实验揭示了MCP在工具描述依赖、文件攻击、链攻击及数据命令区分等方面的漏洞,并提供了安全增强建议。

Comments Accepted by IEEE Transactions on Dependable and Secure Computing (TDSC). $\href{https://ieeexplore.ieee.org/abstract/document/11531012}{Official \ version}$

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24426 2026-05-26 cs.CL 81%

SEAL: Synergistic Co-Evolution of Agents and Learning Environments

SEAL: 智能体与学习环境的协同共演化

Yihao Hu, Zhihao Wen, Xiujin Liu, Pan Wang, Xin Zhang, Wei Wu

机构 * Ant Group(蚂蚁集团) Westlake University(西湖大学) University of Michigan--Ann Arbor(密歇根大学安娜堡分校) University of Science and Technology of China(中国科学技术大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出SEAL框架,通过协同演化智能体策略与训练环境,解决智能体与环境错配问题,在低资源多轮工具使用任务中提升性能并实现正向分布外迁移。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13203 2026-05-26 cs.NI cs.AI 81%

Adversarial Network Imagination: Causal LLMs and Digital Twins for Proactive Telecom Mitigation

对抗性网络想象:因果大语言模型与数字孪生用于主动电信缓解

Vignesh Sriram, Yuqiao Meng, Luoxi Tang, Zhaohan Xi

机构 * Binghamton University(宾夕法尼亚州立大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出对抗性网络想象框架,结合因果大语言模型、知识图谱和数字孪生,主动生成、模拟和评估对抗性网络故障,实现从被动故障排查向预期韧性分析的转变。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.21403 2026-05-21 cs.CL 81%

Quantifying the cross-linguistic effects of syncretism on agreement attraction

量化语言间合成影响对一致性的吸引力

Utku Turk, Eva Neu

机构 * University of Maryland, College Park(马里兰大学学院公园分校) University of Massachusetts, Amherst(马萨诸塞大学阿默斯特分校)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究探讨了合成对一致性吸引力的影响在不同语言中的差异,通过大规模语言模型的 surprisal 和注意力熵来分析四种语言中的表现,揭示了合成如何调节吸引力的机制。

Comments SCiL Conference Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.19988 2026-05-20 cs.SE cs.AI cs.DB cs.PF 81%

A Case for Agentic Tuning: From Documentation to Action in PostgreSQL

为代理调优辩护:从文档到PostgreSQL中的行动

Hongyu Lin, Mingyu Li, Weichen Zhang, Yihang Lou, Mingjie Xing, Yanjun Wu, Haibo Chen

机构 * Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) University of Chinese Academy of Sciences(中国科学院大学) Key Laboratory of System Software (Chinese Academy of Sciences)(中国科学院系统软件重点实验室) Beihang University(北航) Peking University(北京大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI

AI总结 本文提出通过动态行动替代静态文档进行系统调优,引入PerfEvolve工具,利用LLM代理实现版本一致性验证、工作负载特定分析和多参数联合优化,实验表明其在PostgreSQL上比现有文档驱动调优方法提升35.2%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.05974 2026-05-20 cs.CR cs.AI 81%

PragLocker: Protecting Agent Intellectual Property in Untrusted Deployments via Non-Portable Prompts

PragLocker: 通过非可移植提示保护代理知识产权

Qinfeng Li, Yuntai Bao, Jianghui Hu, Wenqi Zhang, Jintao Chen, Huifeng Zhu, Yier Jin, Xuhong Zhang

机构 * Zhejiang University(浙江大学) Management Center, School of Software Technology (Ningbo), Zhejiang University(浙江大学软件学院(宁波)管理中心) University of Science and Technology of China(中国科学技术大学) Chang'an University(长安大学) Washington University in St. Louis(圣路易斯华盛顿大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI

AI总结 针对无信任部署中代理提示易被复制和重用的问题,PragLocker提出了一种保护方案,通过构建语义锚定的混淆提示并注入噪声,有效降低跨LLM可移植性,同时保持目标性能和对抗鲁棒性。

Comments accepted to the 43rd International Conference on Machine Learning (ICML 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18673 2026-05-19 cs.CY cs.CL 81%

Generative AI Advertising as a Problem of Trustworthy Commercial Intervention

生成式AI广告作为可信商业干预问题

Jingyi Qiu, Qiaozhu Mei

机构 * University of Michigan(密歇根大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文探讨生成式AI广告如何通过影响生成过程而非内容放置来改变广告方式,提出基于影响层级的分类体系,并指出当前系统在更深层次商业影响上的可信度问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.17137 2026-05-19 cs.AI 81%

Latent Heuristic Search: Continuous Optimization for Automated Algorithm Design

潜在启发式搜索:为自动化算法设计的连续优化

Cheikh Ahmed, Mahdi Mostajabdaveh, Zirui Zhou

机构 * Huawei Technologies Canada(华为加拿大技术有限公司)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种连续启发式发现框架,通过将离散程序映射到连续嵌入空间,并利用可微代理模型进行梯度优化,以提升自动化算法设计的性能。

Comments Accepted at LION 2026, The Learning and Intelligent Optimization Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15184 2026-05-15 cs.CL 81%

Is Grep All You Need? How Agent Harnesses Reshape Agentic Search

Grep 是否足够?代理如何重塑搜索流程

Sahil Sen, Akhil Kasturi, Elias Lumer, Anmol Gulati, Vamse Kumar Subbiah

机构 * PricewaterhouseCoopers(普华永道)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文通过两个实验比较了grep与向量检索在代理搜索中的效果,发现grep在准确性上优于向量检索,但整体表现仍受代理架构和工具调用方式影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.15172 2026-05-15 cs.CR cs.CL 81%

MetaBackdoor: Exploiting Positional Encoding as a Backdoor Attack Surface in LLMs

MetaBackdoor:利用位置编码作为大语言模型中的后门攻击面

Rui Wen, Mark Russinovich, Andrew Paverd, Jun Sakuma, Ahmed Salem

机构 * Institute of Science Tokyo(东京科学研究院) Microsoft Azure(微软Azure) Microsoft Security Response Center(微软安全响应中心)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出MetaBackdoor,通过利用位置信息作为触发器,而非修改文本内容,实现对大语言模型的后门攻击,展示了长度相关的触发器可激活隐蔽后门,并展示了自我激活场景和与内容后门的组合攻击。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.11996 2026-05-13 cs.AI 81%

BadSKP: Backdoor Attacks on Knowledge Graph-Enhanced LLMs with Soft Prompts

BadSKP: 针对增强知识图谱的大型语言模型的后门攻击

Xiaoting Lyu, Yufei Han, Hangwei Qian, Haoyuan Yu, Xiang Ao, Bin Wang, Chenxu Wang, Xiaobo Ma, Wei Wang

机构 * Ministry of Education Key Lab for Intelligent Networks and Network Security(教育部长智能网络与网络安全重点实验室) Xi’an Jiaotong University(西安交通大学) INRIA(法国国家信息与自动化技术研究院) CFAR, A*STAR(新加坡A*STAR机构) Beijing Key Laboratory of Security and Privacy in Intelligent Transportation(北京智能交通安全与隐私重点实验室) Beijing Jiaotong University(北京交通大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) School of Cyber Engineering, Xi’an University of Electronic Science and Technology(西安电子科技大学网络安全工程学院) Ministry of Education Key Lab for Intelligent Networks and Network Security at Xi’an Jiaotong University(西安交通大学教育部长智能网络与网络安全重点实验室)

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文研究了增强知识图谱的大型语言模型中软提示通道的后门攻击问题,提出BadSKP攻击方法,通过多阶段优化策略有效攻击图到提示接口,实验表明其在冻结和毒化设置下具有高成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09287 2026-05-13 cs.AI 81%

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning

PiCA:基于枢纽的搜索代理强化学习中的信用分配

Dongyi Liu, Yifan Niu, Qinwen Wang, Han Xiao, Jia Li

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 PiCA通过重新定义搜索轨迹为累积搜索进度的序列过程,解决长周期信用分配中的稀疏奖励、孤立信用和分布偏移问题,提升知识密集型问答任务性能。

Comments 21 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.10114 2026-05-12 cs.CL 81%

SkillRAE: Agent Skill-Based Context Compilation for Retrieval-Augmented Execution

SkillRAE: 基于技能的上下文编译用于检索增强执行

Xiangcheng Meng, Shu Wang, Yixiang Fang

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 SkillRAE提出一种两阶段检索增强执行方法,通过构建多级技能图和救援感知紧凑编译,提升技能证据组织效率,实验证明在SkillsBench上比SOTA方法提升11.7%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09891 2026-05-12 cs.IR cs.AI 81%

ArchRAG: Attributed Community-based Hierarchical Retrieval-Augmented Generation

ArchRAG: 基于属性社区的分层检索增强生成

Shu Wang, Yixiang Fang, Yingli Zhou, Xilin Liu, Yuchi Ma

机构 * Microsoft(微软)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 ArchRAG通过引入属性社区和层次聚类方法,提升图数据检索效率与生成准确性,降低token消耗。

Comments Published in Proceedings of the AAAI Conference on Artificial Intelligence, 2026

Journal ref Proceedings of the AAAI Conference on Artificial Intelligence, 40(19), 15868-15876, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06479 2026-05-08 stat.ML cs.LG math.ST stat.TH 81%

Risk-Controlled Post-Processing of Decision Policies

受风险控制的决策策略后处理

Sunay Joshi, Tao Wang, Hamed Hassani, Edgar Dobriban

机构 * University of Pennsylvania(宾夕法尼亚大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.LG

AI总结 本文研究了在风险约束下如何通过后处理优化决策策略,提出了一种基于阈值结构的后处理方法,能够在保持基线策略一致性的前提下控制风险,实验验证了其在医疗诊断、LLM路由和多类决策任务中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24768 2026-05-08 cs.AI 81%

Supervising Ralph Wiggum: Exploring a Metacognitive Co-Regulation Agentic AI Loop for Engineering Design

监督拉尔夫·维格姆:探索一种元认知共调节智能体循环用于工程设计

Zeda Xu, Nikolas Martelaro, Christopher McComb

机构 * Department of Mechanical Engineering(机械工程系) Carnegie Mellon University(卡内基梅隆大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种自我调节循环和共调节设计智能体循环,通过元认知共调节缓解设计固定问题,提升工程设计性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.05833 2026-05-08 cs.AI 81%

On the Role of Language Representations in Auto-Bidding: Findings and Implications

语言表示在自动出价中的作用:发现与启示

Guanyu Zhu, Jining Luan, Hanwen Du, Xinyu Fang, Sibo Xu, Ersheng Ni, Hongji Li, Jincheng Fang, Ronghao Chen, Huacan Wang, Xuanqi Lan, Yongxin Ni, Yiqi Sun, Youhua Li

机构 * City University of Hong Kong(香港城市大学) South China Agricultural University(华南农业大学) University of Electronic Science and Technology of China(电子科技大学) The Ohio State University, Columbus(俄亥俄州立大学) Hefei University of Technology(合肥工业大学) School of Economics and Management, Wuhan University(武汉大学经济管理学院) Faculty of Engineering, The University of Queensland(昆士兰大学工程学院) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) The Hong Kong University of Science and Technology(香港科学大学) Peking University(北京大学) University of Chinese Academy of Sciences(中国科学院大学) Santa Clara University(圣克拉拉大学) Boston University(波士顿大学) The University of Hong Kong(香港大学)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究语言表示在自动出价中的作用,提出SemBid框架,通过整合语义与数值信息提升出价可控性和泛化能力,实验表明其在多种场景下优于现有方法。

Comments 19 page

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20854 2026-05-08 cs.CL cs.IR 81%

How important is Recall for Measuring Retrieval Quality?

召回在衡量检索质量时有多重要?

Shelly Schwartz, Oleg Vasilyev, Randy Sawaya

机构 * Primer Technologies Inc.

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL

AI总结 本文评估了在大型动态知识库中,如何通过LLM判断响应质量来衡量检索质量,提出了一种无需知道总相关文档数的高效方法。

Comments Dataset: https://huggingface.co/datasets/primer-ai/retrieval-response

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00844 2026-05-05 cs.CY cs.AI 81%

The Oracle's Fingerprint: Correlated AI Forecasting Errors and the Limits of Bias Transmission

预言者指纹:相关AI预测误差与偏见传递的极限

Theodor Spiro

机构 * Theodor Spiro

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究揭示大型语言模型预测误差的高度相关性,以及人类群体预测受其影响的机制,发现AI系统与人类偏见模式的相似性及传递效应。

Comments 23 pages, 3 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00827 2026-05-05 cs.DC cs.AI cs.SE 81%

Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol

将智能与执行分离:一种用于模型上下文协议的工作流引擎

Abhinav Singh Parmar

机构 * Infosys

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种MCP原生编排层,将智能决策与执行分离,通过工作流蓝图减少token消耗,提升执行效率。

Comments 16 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00782 2026-05-04 cs.SE cs.AI 81%

GeoContra: From Fluent GIS Code to Verifiable Spatial Analysis with Geography-Grounded Repair

GeoContra:从流畅的GIS代码到可验证的空间分析与地理基础修复

Yinhao Xiao, Rongbo Xiao, Yihan Zhang

机构 * School of Big Data and Artificial Intelligence, Guangdong University of Finance and Economics(大数据与人工智能学院,广东财经大学) School of Geography and Environment Economics, Guangdong University of Finance & Economics(地理与环境经济学院,广东财经大学) Guangdong Engineering Research Center of Low-Altitude Remote Sensing Intelligent Monitoring(低空遥感智能监测工程研究中心)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI

AI总结 GeoContra通过验证和修复框架提升LLM驱动的GIS工作流空间正确性,通过静态检查、运行时验证和语义验证提高空间分析准确性,实验显示在多个模型上空间正确性显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25846 2026-04-29 cs.CR cs.AI 81%

Towards Agentic Investigation of Security Alerts

朝向安全警报的代理调查

Even Eilertsen, Vasileios Mavroeidis, Gudmund Grov

机构 * University of Oslo(奥斯陆大学) Norwegian Defence Research Establishment (FFI)(挪威国防研究机构(FFI))

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种利用大语言模型自动化安全警报初步调查的代理工作流,通过预定义查询和结构化工具访问提高准确性,减少人工工作量。

Comments 10 pages, 3 figures, 4 tables. Accepted at the 2025 IEEE International Conference on Big Data (BigData)

Journal ref Proc. 2025 IEEE Int. Conf. on Big Data (BigData), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17060 2026-04-27 cs.CY cs.AI 81%

Initial results of the Digital Consciousness Model

数字意识模型的初步结果

Derek Shiller, Laura Duffy, Arvo Muñoz Morán, Adrià Moret, Chris Percy, Hayley Clatterbuck

机构 * University of Barcelona(巴塞罗那大学) Co-Sentience Initiative(共意识计划)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI

AI总结 本文探讨了数字意识模型在评估AI系统意识证据方面的初步发现,指出2024年LLM的意识证据不充分,但比更简单AI系统的证据弱。

Comments v1.1 Revised section 4.2 details and acknowledgments

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.06922 2026-04-22 cs.CR cs.LG 81%

Whispers in the Machine: Confidentiality in Agentic Systems

机器中的低语:代理系统中的保密性

Jonathan Evertz, Merlin Chlosta, Lea Schönherr, Thorsten Eisenhofer

机构 * CISPA Helmholtz Center for Information Security(信息安全勒内希特中心)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究探讨了基于大语言模型的代理系统中的保密性问题,通过抽象敏感数据为秘密字符串,评估了十种代理在20种工具场景和14种攻击策略下的安全性,发现所有代理均存在至少一种漏洞,现有防御措施无法有效防止数据泄露。

Comments Accepted at Conference on Detection of Intrusions and Malware & Vulnerability Assessment (DIMVA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏