arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Machine Learning · 会议 · Machine Learning

2026-06-25 至 2026-06-25 共收录 28
2606.25743 2026-06-25 cs.LG 新提交

Black-Box Assisted Regression: Phase Transitions and Minimax Optimality

黑盒辅助回归:相变与极小化最优性

Yan Zhou

机构 * School of Mathematics and Statistics(数学与统计学学院)

AI总结 研究有限标注下利用固定黑盒预测器进行非参数回归的问题,发现风险在临界半径处发生相变,并提出安全残差估计器以避免负迁移,达到极小化最优。

Comments 23 pages, 3 figures. Accepted at the 43rd International Conference on Machine Learning (ICML 2026)

Journal ref Proceedings of the 43rd International Conference on Machine Learning, PMLR 306, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25546 2026-06-25 cs.CV 新提交

Disease-Centric Vision-Language Pretraining with Hybrid Visual Encoding for 3D Computed Tomography

面向3D计算机断层扫描的疾病中心视觉语言预训练与混合视觉编码

Bowen Shi, Weiwei Cao, Ruifeng Yuan, Wanxing Chang, Wenrui Dai, Hongkai Xiong, Ling Zhang, Jianpeng Zhang

机构 * DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) Hupan Lab, 310023, Hangzhou, China(虎扑实验室,杭州,中国) Shanghai Jiao Tong University, China(上海交通大学,中国) Zhejiang University, China(浙江大学,中国) Fudan University, China(复旦大学,中国)

AI总结 提出一种结合CNN-ViT混合编码器、疾病级对比学习和诊断感知提示的视觉语言预训练框架,在CT-RATE和Rad-ChestCT上取得最优性能,并提升零样本诊断可靠性。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25402 2026-06-25 cs.SE cs.AI 新提交

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models

LibEvoBench:探测代码生成模型中的时间知识分层

Daniele Cipollone, Sergey Titov, Maliheh Izadi, Egor Bogomolov, Arie van Deursen

机构 * Faculty of EEMCS, Delft University of Technology, Delft, Netherlands(代尔夫特理工大学电子工程与信息科学学院) JetBrains Research, Amsterdam, Netherlands(JetBrains研究)

AI总结 针对LLM在代码生成中因训练数据时间混合导致API版本混淆的问题,提出多版本基准LibEvoBench和新指标SEUS,揭示模型对版本不敏感且仅靠文档可提升准确性。

Comments Accepted at the DL4Code workshop at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25394 2026-06-25 cs.LG cs.AI 新提交

FactorLibrary: From Polynomials to Circuits via Recursive Subgoals

FactorLibrary: 从多项式到电路通过递归子目标

Rohan Pandey, Michael Ruofan Zeng, Weikun K. Zhang, Kaijie Jin, Naomi Morato, Archit Ganapule, Bhaumik Mehta, Jarod Alper

机构 * University of Washington, Seattle, WA, USA(华盛顿大学)

AI总结 将有限域上多项式的最小算术电路发现建模为强化学习问题,提出FactorLibrary存储可分解子表达式作为子目标,使用PPO+MCTS的top-down方法在复杂度8以内达到91.8%的成功率。

Comments 14 pages, 8 figures, in 3rd AI for Math Workshop (ICML 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25389 2026-06-25 cs.AI 新提交

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse

基于技能划分与复用的离线多智能体持续协作

Yuchen Xiao, Lei Yuan, Ruiqi Xue, Tieyue Yin, Yang Yu

机构 * National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室) School of Artificial Intelligence, Nanjing University(南京大学人工智能学院) Polixir Technologies.(Polixir技术公司) Kuang Yaming Honors School, Nanjing University(匡杨明荣誉学院,南京大学)

AI总结 提出COMAD框架,通过自编码器从离线数据中发现可复用技能,并利用密度估计器指导优势函数,解决多智能体持续学习中的灾难性遗忘和干扰问题。

Comments 29 pages, 12 figures, ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25318 2026-06-25 cs.CV cs.LG 新提交

REViT: Roto-reflection Equivariant Convolutional Vision Transformer

REViT: 旋转反射等变卷积视觉Transformer

Sheir A. Zaheer, Alexander C. Holston, Chan Y. Park

机构 * KC Machine Learning Lab(韩国首尔计算机机器学习实验室)

AI总结 提出一种离散旋转反射群等变的视觉Transformer,通过卷积注意力机制保持特征图的旋转、翻转和位置对称性,在图像分类任务中优于现有方法。

Comments Accepted for publication at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25273 2026-06-25 cs.CV 新提交

CoGeoAD: Hierarchical Color-Geometric Fusion with Multi-View Attention for Zero-Shot 3D Anomaly Detection

CoGeoAD: 基于层次化颜色-几何融合与多视角注意力的零样本3D异常检测

Ke Xu, Xinle Wang, Yanning Hou, Xueliang Ma, Juan Xie, Jianfeng Qiu

机构 * State Key Laboratory of Opto-Electronic Information Acquisition Protection Technology, Anhui University, Hefei, China School of Artificial Intelligence, Anhui University, Hefei, China College of Intelligence Science Technology, National University of Defense Technology, Changsha, China School of Mathematics \& Physics, Anhui Jianzhu University, Hefei, China

AI总结 提出CoGeoAD框架,通过像素对齐的多视图图像融合颜色与几何特征,利用数据驱动的多视角注意力和多阶段颜色-几何融合模块,在MVTec3D-AD和Eyecandies基准上实现零样本3D异常检测的最新性能。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25182 2026-06-25 cs.CL cs.AI cs.LG 新提交

What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics

中间层知道什么:从熵动力学检测越狱

Sofiia Nikolenko, Michele Papucci, Mina Rezaei, Shireen Kudukkil Manchingal

机构 * LMU Munich(慕尼黑大学) relAI – Konrad Zuse School of Excellence in Reliable AI(relAI – 康拉德·楚泽可靠人工智能卓越学校) University of Pisa(比萨大学) Munich Center for Machine Learning(慕尼黑机器学习中心) School of Engineering, Computing and Mathematics, Oxford Brookes University(牛津布鲁克斯大学工程、计算与数学学院)

AI总结 通过分析冻结LLM各层的token级预测熵轨迹,发现中间层的熵动力学特征(如基于排名的单调趋势分数)能有效检测越狱攻击,且无需额外训练。

Comments Accepted at the European Conference on Machine Learning and Principles and Practice of Knowledge Discovery in Databases (ECML PKDD) 2026. A short version accepted at EIML@ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24993 2026-06-25 cs.LG 新提交

The Geometry of Sequential Learning: Lie-Bracket Prediction of Transfer Order

序列学习的几何:李括号预测迁移顺序

John Sweeney

机构 * Sideplane AI

AI总结 提出李括号对偶性分数预测序列学习中的源域顺序,通过李括号锦标赛实现O(N log N)排序,在指令微调和DPO中达到高准确率。

Comments Accepted to ICML 2026. 20 pages, including appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24957 2026-06-25 cs.CL cs.LG 新提交

Dustin: Draft-Augmented Sparse Verification for Efficient Long-Context Generation with Speculative Decoding

Dustin: 用于高效长上下文推测解码的草稿增强稀疏验证

WenHung Lee, Jian-Jia Chen, Xiaolin Lin, Pei-Shuo Wang, Chi-Chih Chang, Chun-Che Yang, Ning-Chi Huang, Grace Li Zhang, Kai-Chiang Wu

机构 * National Yang Ming Chiao Tung University(国立阳明交通大学) Cornell University(康奈尔大学)

AI总结 提出Dustin框架,利用草稿模型的前瞻信号和目标模型的历史注意力,在推测解码中实现高效稀疏验证,显著加速长上下文生成。

Comments Accepted to ICML 2026. 9 pages main text, includes references and appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25761 2026-06-25 cs.LG math.OC 新提交

Bridging Spherical Black-Box Optimizers

桥接球形黑箱优化器

Johannes Ackermann, Stefano Peluchetti

机构 * The University of Tokyo(东京大学)

AI总结 将进化策略、共识优化和积分优化统一为理论框架,通过引入混合优化器控制平坦偏好和模态,在连续控制和语言模型合并任务中提升性能与鲁棒性。

Comments Accepted for publication at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26050 2026-06-25 cs.LG cond-mat.dis-nn cs.AI cs.CL 新提交

Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining

自然去突现:不对称控制哪些规则在预训练中幸存

Juliana Li, Diya Sreedhar

机构 * Harvard University(哈佛大学)

AI总结 发现语言模型在预训练中学习规则后会自动遗忘,规则存亡由训练数据中支持频率决定,且遗忘不可逆。

Comments Foundations of Deep Generative Models (FoGen) Workshop at ICML 2026. 23 pages (5-page main text plus appendices), 5 figures. Code: https://github.com/lijuliana/Natural-Ungrokking

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.05097 2026-06-25 cs.LG cs.AI cs.CL 版本更新

Continual Knowledge Updating in LLM Systems: Learning Through Multi-Timescale Memory Dynamics

连续知识更新在大语言模型系统中:通过多时间尺度记忆动态学习

Andreas Pattichis, Constantine Dovrolis

机构 * The Cyprus Institute, Nicosia, Cyprus(塞浦路斯研究所,尼科西亚,塞浦路斯)

AI总结 本文提出基于多时间尺度动态的记忆机制,通过耦合内部变量实现知识的持续更新与学习,重构外部记忆作为学习的基础 substrates。

Comments Accepted as a poster at the ICML 2026 Workshop "Continual Adaptation at Scale: Towards Sustainable AI" (CATS@ICML 2026). 9 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25496 2026-06-25 cs.AI 版本更新

Improving Zero-Shot Offline RL via Behavioral Task Sampling

通过行为任务采样改进零样本离线强化学习

Nazim Bendib, Nicolas Perrin-Gilbert, Olivier Sigaud

AI总结 本文提出通过直接从离线数据集中提取任务向量来改进零样本离线强化学习,提升任务泛化能力,实验表明在多个基准环境中性能提升20%。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.07904 2026-06-25 cs.LG cs.CV cs.NE 版本更新

Kuramoto Oscillatory Phase Encoding: Neuro-inspired Synchronization for Improved Learning Efficiency

Kuramoto振荡相位编码:神经启发同步机制提升学习效率

Mingqing Xiao, Yansen Wang, Dongqi Han, Caihua Shan, Dongsheng Li

机构 * Microsoft Research Asia(微软亚洲研究院)

AI总结 提出Kuramoto振荡相位编码(KoPE),为视觉Transformer引入神经启发同步机制,通过增强结构学习提升训练、参数和数据效率,并在语义分割、少样本推理等任务中取得优势。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06566 2026-06-25 cs.CV cs.AI cs.CL 版本更新

SPARC: Separating Perception And Reasoning Circuits for Test-time Scaling of VLMs

SPARC: 分离感知与推理回路以实现VLM的测试时扩展

Niccolo Avogaro, Nayanika Debnath, Li Mi, Thomas Frick, Junling Wang, Zexue He, Hang Hua, Konrad Schindler, Mattia Rigotti

机构 * ETH Zürich(苏黎世联邦理工学院) IBM Research(IBM研究院) MIT-IBM Watson AI Lab(麻省理工-IBM沃森人工智能实验室)

AI总结 提出SPARC框架,通过显式分离视觉感知与推理,实现测试时动态扩展和不对称计算分配,在视觉推理任务中优于基线方法。

Comments Accepted at the 43rd International Conference on Machine Learning (ICML 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05143 2026-06-25 cs.AI cs.IR 版本更新

CausalRAG2: Hierarchical Causal Knowledge Graph Design for RAG

CausalRAG2: 面向RAG的分层因果知识图谱设计

Nengbo Wang, Tuo Liang, Vikash Singh, Chaoda Song, Van Yang, Yu Yin, Jing Ma, Jagdip Singh, Vipin Chaudhary

机构 * Department of Computer and Data Sciences, Case Western Reserve University, Cleveland, OH, USA(计算机与数据科学系,凯斯西储大学,克利夫兰,OH,USA) Design and Innovation Department, Case Western Reserve University, Cleveland, OH, USA(设计与创新部门,凯斯西储大学,克利夫兰,OH,USA)

AI总结 提出CausalRAG2框架,通过分层模块间的因果门控机制显式建模因果关系,抑制虚假相关,实现大规模知识图谱上的可扩展推理,并引入HolisQA基准,实验表明其优于现有图基RAG方法。

Comments Accepted at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02877 2026-06-25 cs.LG 版本更新

A Geometry-Aware Efficient Algorithm for Compositional Entropic Risk Minimization

组合熵风险最小化的几何感知高效算法

Xiyuan Wei, Linli Zhou, Bokun Wang, Chih-Jen Lin, Tianbao Yang

机构 * Texas A\&M University(德克萨斯A&M大学) University of Texas, Austin(德克萨斯大学奥斯汀分校) National Taiwan University(国立台湾大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

AI总结 针对组合熵风险最小化问题,提出几何感知随机算法SCENT,利用负指数函数诱导的Bregman散度进行随机近端镜像下降,实现凸问题O(1/√T)收敛率,并在极端分类、部分AUC最大化等任务中优于基线。

Comments Accepted to 43rd International Conference on Machine Learning. 39 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01903 2026-06-25 cs.LG stat.ML 版本更新

Data- and Variance-dependent Regret Bounds for Online Tabular MDPs

在线表格MDPs的数据依赖和方差依赖遗憾界

Mingyi Li, Taira Tsuchiya, Kenji Yamanishi

机构 * Department of Mathematical Informatics, The University of Tokyo, Tokyo, Japan(数学信息学系,东京大学,东京,日本)

AI总结 针对已知转移的在线表格马尔可夫决策过程,提出在对抗性环境下实现数据依赖遗憾界、在随机环境下实现方差依赖遗憾界的最优算法,并证明全局优化方法达到近乎最优。

Comments Accepted at ICML 2026. 72 pages, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.23075 2026-06-25 cs.LG cs.RO 版本更新

RN-D: Discretized Categorical Actors for On-Policy Reinforcement Learning

RN-D:用于同策略强化学习的离散化分类演员

Yuexin Bian, Jie Feng, Tao Wang, Yijiang Li, Sicun Gao, Yuanyuan Shi

机构 * University of California San Diego, La Jolla, USA(加州大学圣地亚哥分校)

AI总结 提出离散化分类演员替代高斯演员,结合正则化网络形成RN-D,在同策略RL中实现连续控制的最优性能。

Comments Accepted by ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.17037 2026-06-25 cs.CV cs.AI 版本更新

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs

AMVICC: 一种用于VLM和IGM跨模态故障模式分析的新型基准

Aahana Basappa, Pranay Goel, Anusri Karra, Anish Karra, Asa Gilmore, Kevin Zhu

机构 * Centennial High School, Frisco, Texas, USA(Centennial High School, Texas, USA) Lebanon Trail High School, Frisco, Texas, USA(Lebanon Trail High School, Texas, USA) West Windsor-Plainsboro High School, Princeton Junction, New Jersey, USA(West Windsor-Plainsboro High School, New Jersey, USA) Algoverse AI Research, Palo Alto, California, USA(Algoververse AI Research, California, USA)

AI总结 提出AMVICC基准,通过图像到文本和文本到图像任务系统比较多模态大模型和图像生成模型的视觉推理失败模式,发现故障模式在模型和模态间共享,但存在特定于模型和模态的失败。

Comments 14 pages, 4 figures, 8 tables. Presented at the 39th Conference on Neural Information Processing Systems Workshop: VLM4RWD. Presented at the 43th International Conference on Machine Learning Workshops: ICML 2026 CTB, ICML 2026 FAGEN, ICML 2026 EMM-QA. Authors Aahana Basappa and Pranay Goel contributed equally. Code: https://github.com/AahanaB24/AMVICC, Data: https://doi.org/10.5281/zenodo.17646068

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26376 2026-06-25 cs.CV 版本更新

ScalingAR: Scaling Confidence for Autoregressive Image Generation

ScalingAR: 自回归图像生成的置信度缩放

Harold Haodong Chen, Xianfeng Wu, Wen-Jie Shu, Rongjin Guo, Disen Lan, Harry Yang, Ying-Cong Chen

AI总结 提出ScalingAR框架,通过令牌熵作为置信度信号,在轮廓级和策略级进行自适应轨迹剪枝和动态引导调度,无需早期解码或外部奖励,显著提升自回归图像生成性能。

Comments ICML 2026; Code: https://github.com/EnVision-Research/ScalingAR

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00866 2026-06-25 cs.LG cs.CL 版本更新

Removing Noise, not Finding Gold: Quality Filtering for Large-Scale Pretraining

去除噪声,而非寻找黄金:大规模预训练的质量过滤

Thiziri Nait Saada, Louis Bethune, Michal Klein, David Grangier, Marco Cuturi, Pierre Ablin

机构 * Work done as an intern at Apple(苹果公司实习生) University of Oxford(牛津大学) Apple(苹果公司)

AI总结 本文深入分析基于分类器的质量过滤(CQF)方法,发现其虽提升下游任务性能,但会隐式过滤高质量数据,且与基于合成数据的模型行为趋势不同,质疑CQF捕捉数据质量的有效性。

Comments 21 pages, 22 figures, 2 tables, accepted at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01163 2026-06-25 cs.LG stat.ML 版本更新

How Does the Pretraining Distribution Shape In-Context Learning? A Fundamental Trade-Off

预训练分布如何塑造上下文学习?一个基本的权衡

Waïss Azizian, Ali Hasan

AI总结 通过理论框架和实验,揭示预训练分布的统计特性(如重尾行为)在上下文学习中导致任务选择鲁棒性与泛化能力之间的基本权衡。

Comments 57 pages, 15 figures; to be presented at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19532 2026-06-25 cs.LG 版本更新

Fox in the Henhouse: Supply-Chain Backdoor Attacks Against Reinforcement Learning

狐狸进鸡窝:针对强化学习的供应链后门攻击

Shijie Liu, Andrew C. Cullen, Paul Montague, Sarah Erfani, Benjamin I. P. Rubinstein

机构 * School of Computing and Information Systems, University of Melbourne, Parkville, Australia(墨尔本大学计算机与信息系统学院,墨尔本,澳大利亚)

AI总结 提出SCAB攻击,通过污染仅3%的训练经验即可激活超过90%的触发动作,使受害者平均每轮回报降低80%,揭示了强化学习供应链中的后门攻击风险。

Comments Forty-Third International Conference on Machine Learning (ICML2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.02205 2026-06-25 cs.LG 版本更新

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control

从不确定到安全:扩散模型的安全自适应用于PDE控制

Peiyan Hu, Xiaowei Qian, Wenhao Deng, Rui Wang, Haodong Feng, Ruiqi Feng, Tao Zhang, Long Wei, Yue Wang, Zhi-Ming Ma, Tailin Wu

机构 * intern(实习生) Westlake University(西湖大学) School of Engineering, Westlake University(西湖大学工程学院) Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院) Fudan University(复旦大学) Zhongguancun Academy(中关村学院)

AI总结 提出SafeDiffCon方法,通过共形预测估计不确定性分位数,结合后训练和推理阶段调整扩散模型,在满足安全约束下实现最优PDE控制。

Comments ICML 2025. 24 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03393 2026-06-25 cs.LG

Prediction Models That Learn to Avoid Missing Values

能够避免缺失值的学习预测模型

Lena Stempfle, Anton Matsson, Newton Mwai, Fredrik D. Johansson

机构 * Department of Computer Science and Engineering, Chalmers University of Technology and University of Gothenburg(计算机科学与工程系,查尔姆斯理工大学和哥德堡大学) Chalmers University of Technology(查尔姆斯理工大学) University of Gothenburg(哥德堡大学)

AI总结 本文提出一种避免缺失值的机器学习框架,通过在决策树、树集成和稀疏线性模型中引入特定分类器的正则化项,减少测试时对缺失值的依赖,同时保持预测性能。

Journal ref Proceedings of the 42nd International Conference on Machine Learning, PMLR 234, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.07239 2026-06-25 cs.LG stat.ML 版本更新

RotRNN: Modelling Long Sequences with Rotations

RotRNN:利用旋转建模长序列

Kai Biegun, Rares Dolga, Jake Cunningham, David Barber

机构 * UCL AI Centre(UCL人工智能中心) UiPath London, UK(UiPath伦敦英国)

AI总结 提出RotRNN线性循环模型,利用旋转矩阵简化初始化与归一化,在长序列基准上达到竞争性能。

Comments Next Generation of Sequence Modeling Architectures Workshop at ICML 2024

详情

展开后加载摘要…

URL PDF HTML 收藏