arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Georgia Institute of Technology(佐治亚理工学院)

2026-07-14 至 2026-07-14 共收录 4
2606.17034 2026-07-14 cs.CL cs.LG 版本更新

KVEraser: Learning to Steer KV Cache for Efficient Localized Context Erasing

KVEraser: 学习操控KV缓存以实现高效的局部上下文擦除

Mufei Li, Shikun Liu, Dongqi Fu, Haoyu Wang, Yinglong Xia, Hong Li, Hong Yan, Pan Li

机构 * Georgia Institute of Technology(佐治亚理工学院) Meta

AI总结 提出KVEraser方法,通过学习操控KV缓存实现局部上下文擦除,避免全局重计算,在长上下文任务中接近全重算性能且延迟仅增加24%。

Comments Oral at the ICML 2026 Workshop on the Impact of Memorization on Trustworthy Foundation Models; Code available at https://github.com/Graph-COM/KVEraser

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.19029 2026-07-14 cs.RO 版本更新

Distributionally Robust Control via Stein Variational Inference for Contact-Rich Manipulation

通过Stein变分推断进行分布鲁棒控制的接触丰富操作

Hrishikesh Sathyanarayan, Victor Vantilborgh, Harish Ravichandar, Tom Lefebvre, Ian Abraham

机构 * Department of Mechanical Engineering, Yale University(耶鲁大学机械工程系) Department of Electromechanical, Systems and Metal Engineering, Ghent University(根特大学机电系统与金属工程系) School of Interactive Computing, Georgia Institute of Technology(佐治亚理工学院交互计算学院) Department of Electrical Engineering, University of Sydney(悉尼大学电子工程系)

AI总结 本文提出了一种基于Stein变分推断的分布鲁棒控制方法,用于提升接触丰富操作中的不确定性建模能力,通过更灵活的不确定性建模在保持性能的同时精确适应不确定性,实验结果表明在广泛参数不确定性下,鲁棒性提高了3倍。

Comments In Proceedings of Robotics: Science and Systems, Sydney, Australia, July 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01331 2026-07-14 cs.CL cs.AI cs.LG 版本更新

MetaState: Persistent Working Memory Enhances Reasoning in Discrete Diffusion Language Models

MetaState: 持久工作记忆增强离散扩散语言模型的推理能力

Kejing Xia, Mingzhe Li, Lixuan Wei, Zhenbang Du, Xiangchi Yuan, Dachuan Shi, Qirui Jin, Wenke Lee

机构 * Georgia Institute of Technology(佐治亚理工学院) University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校) Harvard University(哈佛大学)

AI总结 MetaState通过引入轻量级循环增强模块,为冻结的离散扩散语言模型提供持久固定大小的工作记忆,提升推理性能,平均提升4.5%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08233 2026-07-14 cs.LG 版本更新

Enhancing Reasoning for Diffusion LLMs via Distribution Matching Policy Optimization

通过分布匹配策略优化增强扩散大语言模型的推理能力

Yuchen Zhu, Wei Guo, Jaemoo Choi, Petr Molodyk, Bo Yuan, Molei Tao, Yongxin Chen

机构 * Georgia Institute of Technology(佐治亚理工学院)

AI总结 本文提出DMPO方法,通过分布匹配优化提升扩散大语言模型的推理能力,实现显著的准确率提升。

详情

展开后加载摘要…

URL PDF HTML 收藏