arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Cambridge(剑桥大学)

共收录 1283
2605.29184 2026-05-29 cs.LG cs.AI

Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search with Granular Feedback

影响引导的符号回归:基于大语言模型与细粒度反馈的方程搜索科学发现

Evgeny S. Saveliev, Samuel Holt, Nabeel Seedat, David L. Bentley, Jim Weatherall, Mihaela van der Schaar

机构 * University of Cambridge(剑桥大学) Thomson Reuters Foundational Research(汤姆森·路透基础研究) U. Colorado, Anschutz Medical Campus(科罗拉多大学安舒茨医疗校区)

AI总结 提出影响引导符号回归(IGSR)方法,利用大语言模型生成候选函数并通过细粒度影响分数进行剪枝,结合蒙特卡洛树搜索高效探索组合空间,在多个基准和真实生物数据中发现新关系。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28802 2026-05-28 cs.CL

Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization

人类标注变异作为稳定信号:通过跨标注者偏好优化学习标注者特定的解释行为

Beiduo Chen, Pingjun Hong, Ziyun Zhang, Benjamin Roth, Anna Korhonen, Barbara Plank

机构 * MaiNLP Center for Information and Language Processing(信息与语言处理中心) LMU Munich(慕尼黑大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) University of Vienna(维也纳大学) LTL University of Cambridge(剑桥大学)

AI总结 研究大语言模型能否学习并复现标注者特定的标签-解释行为,提出跨标注者偏好优化(CAPO)方法,通过对比目标标注者与其他有效但非目标标注者的响应来提升模仿和归因能力。

Comments 43 pages, 20 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28253 2026-05-28 cs.CL cs.DB cs.HC

Building Community-Centred NLP Resources for Puno Quechua

构建以社区为中心的普诺克丘亚语自然语言处理资源

Elwin Huaman, Adrian Gamarra Lafuente, Johanna Cordova, Anna Korhonen

机构 * University of Cambridge (UK)(剑桥大学(英国)) Stanford University (USA)(斯坦福大学(美国)) ERTIM - Inalco (France)(ERTIM - Inalco(法国))

AI总结 通过参与式设计收集66小时语音数据,微调Whisper-base等模型,首次为普诺克丘亚语建立ASR基准并开源所有资源。

Comments Sixth Workshop on NLP for Indigenous Languages of the Americas (AmericasNLP 2026), co-located with ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28153 2026-05-28 physics.ao-ph cs.LG

Skillful high-resolution weather forecasting independent of physical models

独立于物理模型的高分辨率天气预报

Pengcheng Zhao, Siqi Xiang, Weixin Jin, Zekun Ni, Jiang Bian, Zuliang Fang, Hongyu Sun, Bin Zhang, Richard E. Turner, Jonathan Weyn, Haiyu Dong, Kit Thambiratnam, Qi Zhang

机构 * Microsoft Corporation(微软公司) University of Cambridge(剑桥大学) The Alan Turing Institute(艾伦·图灵研究所)

AI总结 提出ObsCast系统,仅使用观测数据训练,无需数值天气预报数据,实现短期高分辨率区域预报,性能优于传统NWP。

Comments 26 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27734 2026-05-28 cs.LG

Learn from your own latents and not from tokens: A sample-complexity theory

从自身潜在表示而非token学习:样本复杂度理论

Daniel J. Korchinski, Alessandro Favero, Matthieu Wyart

机构 * Institute of Physics(物理研究所) University of Cambridge(剑桥大学) Johns Hopkins University(约翰霍普金斯大学) EPFL(苏黎世联邦理工学院)

AI总结 本文通过概率上下文无关语法数据,证明潜在预测方法在样本复杂度上相比token级SSL具有指数级优势,并分析了多尺度层次结构的必要性。

Comments 10 pages, 5 figures in main. 28 pages, 14 figures, 1 table in all

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27643 2026-05-28 cs.RO physics.optics

Agentic Language-to-Objective Synthesis for Optofluidic Assembly

面向光流组件的智能语言到目标合成

Ivan Saraev, Elena Erben, Weida Liao, Fan Nan, Gerhard Neumann, Eric Lauga, Moritz Kreysing

机构 * Institute of Biological and Chemical Systems, Karlsruhe Institute of Technology, Germany(马克斯·普朗克研究所生物和化学系统研究所,卡尔斯鲁厄技术大学,德国) Department of Applied Mathematics and Theoretical Physics, University of Cambridge, UK(应用数学和理论物理系,剑桥大学,英国) Department of Mathematics, Imperial College London, UK(数学系,伦敦帝国理工学院,英国) Institute of Anthropomatics and Robotics (IAR), Karlsruhe Institute of Technology, Germany(人机学与机器人研究所(IAR),卡尔斯鲁厄技术大学,德国)

AI总结 提出Speak-to-Objective模块化智能流水线,利用条件大语言模型将口语或书面指令转换为可微目标函数,实现光流控微粒子组装,并支持用户反馈学习。

Comments 21 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20665 2026-05-28 cs.AI

The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models

推理的形状:大型语言模型中推理轨迹的拓扑分析

Xue Wen Tan, Nathaniel Tan, Galen Lee, Stanley Kok

机构 * University of Cambridge, Department of Engineering, England(剑桥大学工程系) National University of Singapore, School of Computing, Singapore(新加坡国立大学计算机学院)

AI总结 提出基于拓扑数据分析(TDA)的评估框架,通过捕捉推理轨迹的几何结构实现高效自动评估,实验表明拓扑特征比图指标更有效预测推理质量。

Comments Accepted in ICML 2026 Workshop: Epistemic Intelligence in Machine Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.12515 2026-05-28 cs.CL

Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation

通过共识驱动的偏好优化缓解多语言大模型中的跨语言文化不一致性

Lucas Resck, Isabelle Augenstein, Anna Korhonen

机构 * Language Technology Lab, University of Cambridge(剑桥大学语言技术实验室) University of Copenhagen(哥本哈根大学)

AI总结 提出C-3PO框架,通过共识驱动的偏好优化,缓解多语言大模型在用户身份明确时因提示语言变化导致的跨语言文化不一致问题,显著提升一致性指标κ_S。

Comments 24 pages, 13 figures, 11 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18647 2026-05-28 cs.LG cs.AI cs.CV cs.IT math.IT

Noise Scheduling as Information-Guided Allocation in Diffusion Training

噪声调度作为扩散训练中的信息引导分配

Gabriel Raya, Bac Nguyen, Georgios Batzolis, Yuhta Takida, Dejan Stancevic, Naoki Murata, Chieh-Hsin Lai, Yuki Mitsufuji, Luca Ambrogioni

机构 * Tilburg University & JADS(蒂尔堡大学及JADS) Sony AI(索尼人工智能) University of Cambridge(剑桥大学) Radboud University(拉德堡德大学) Sony Group Corporation(索尼集团公司)

AI总结 提出InfoNoise,一种在线自适应噪声调度方法,通过估计条件熵率剖面动态调整训练噪声分布,以优化去噪任务中的信息增益,在图像、DNA和语言生成等任务中达到或超越基线,并节省高达3倍训练计算量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22554 2026-05-28 cs.LG

DeepC4: Deep Conditional Census-Constrained Clustering for Large-scale Multitask Spatial Disaggregation of Urban Morphology

DeepC4: 用于城市形态大规模多任务空间分解的深度条件普查约束聚类

Joshua Dimasaka, Christian Geiß, Emily So

机构 * Department of Architecture, University of Cambridge(剑桥大学建筑系) Cambridge University Centre for Risk in the Built Environment(剑桥大学建筑环境风险研究中心) Earth Observation Center, German Aerospace Center(德国航天中心地球观测中心) Institute of Geography, University of Bonn(波恩大学地理研究所)

AI总结 提出DeepC4,一种结合局部普查统计作为聚类约束并联合学习卫星图像模式的多任务深度学习方法,用于城市形态的粗到细空间分解,在卢旺达数据上优于现有方法。

Comments Major Revised Preprint Submitted to ISPRS Journal of Photogrammetry and Remote Sensing (in review) | Keywords: urban morphology, building exposure, physical vulnerability, spatial disaggregation, deep clustering | Data: https://doi.org/10.5281/zenodo.13119552 | Code: https://github.com/riskaudit/DeepC4

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04144 2026-05-28 cs.CV cs.GR

Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation

Chirpy3D: 面向创意细粒度物体生成的部件感知多视角扩散

Kam Woh Ng, Jing Yang, Jia Wei Sii, Chee Seng Chan, Jiankang Deng, Yi-Zhe Song, Tao Xiang, Xiatian Zhu

机构 * University of Surrey(萨里大学) University of Cambridge(剑桥大学) Universiti Malaya(马来亚大学) Imperial College London(伦敦帝国学院)

AI总结 提出Chirpy3D,一种部件感知多视角扩散框架,从无姿态2D图像中学习层次化部件潜在空间,实现部件级交换、插值和零样本组合,无需3D数据或手动标注。

Comments 20 pages. Code at https://github.com/kamwoh/chirpy3d

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27073 2026-05-27 cs.LG

Learning to Orchestrate Agents under Uncertainty

学习在不确定性下编排智能体

Mary Chriselda Antony Oliver, Lan Jiang, Aaron Bundi Anampiu, Elaf Almahmoud, Francesco Quinzan, Umang Bhatt

机构 * Department of Applied Mathematics and Theoretical Physics, University of Cambridge(应用数学与理论物理系,剑桥大学) Centre for Human-Inspired Artificial Intelligence, University of Cambridge(启发式人工智能中心,剑桥大学) African Institute for Mathematical Sciences, South Africa(南非数学科学研究所) Department of Engineering Science, University of Oxford(工程科学系,牛津大学)

AI总结 提出BOT-Orch框架,将编排问题转化为带正则化的多臂赌博机问题,在不确定性下实现异构智能体的自适应编排,理论保证遗憾界为O(√T)并优于基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27068 2026-05-27 cs.CL cs.AI cs.MA

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

QUACK: 多模态社交推理智能体中的沟通知识质疑、理解与审计

Ye Yuan, Rui Song, Weien Li, Zeyu Li, Haochen Liu, Xiangyu Kong, Changjiang Han, Yonghan Yang, Zichen Zhao, Zixuan Dong, Fuyuan Lyu, Bowei He, Haolun Wu, Jikun Kang, Xue Liu

机构 * McGill University(麦吉尔大学) Mila - Quebec AI Institute(魁北克人工智能研究所) University of Cambridge(剑桥大学) MBZUAI - Mohamed bin Zayed University of Artificial Intelligence(MBZUAI - 摩苏尔·本·扎耶德人工智能大学) University of Toronto(多伦多大学) Salesforce

AI总结 提出QUACK框架,通过游戏结果、行为轨迹和话语一致性三级评估,自动审计多模态社交推理智能体语言与感知行为的一致性,发现最强智能体仍有15.1%的空间幻觉和过半无据指控。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26850 2026-05-27 cs.LG

Learning Energy-Based Models from Stochastic Interpolants using Spatiotemporal Differences

从随机插值中学习基于能量的模型:利用时空差异

Hanlin Yu, RuiKang OuYang, Partha Kaushik, Arto Klami, Michael U. Gutmann, Omar Chehab

机构 * University of Helsinki(赫尔辛基大学) University of Cambridge(剑桥大学) Carnegie Mellon University(卡内基梅隆大学) University of Edinburgh(爱丁堡大学)

AI总结 提出时空噪声对比估计(stNCE)框架,通过联合时空差异从随机插值中学习能量函数,统一现有方法并实现与最先进密度估计方法竞争的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26835 2026-05-27 cs.AI

Helicase: Uncertainty-Guided Supply Chain Knowledge Graph Construction with Autonomous Multi-Agent LLMs

Helicase: 不确定性引导的供应链知识图谱构建与自主多智能体大语言模型

Yunbo Long, Haolang Zhao, Ge Zheng, Alexandra Brintrup

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) The Alan Turing Institute(艾伦·图灵研究所)

AI总结 提出Helicase,一种基于多智能体大语言模型的自主系统,通过不确定性引导的迭代验证和知识图谱构建,解决供应链中需要多跳推理的结构化推断问题,并引入SCQA基准评估。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26823 2026-05-27 cs.CL

Generating Logically Consistent Synthetic Supply Chain Data with LLM-Driven Knowledge Graph Reasoning

基于LLM驱动知识图谱推理生成逻辑一致的合成供应链数据

Yunbo Long, Ge Zheng, Liming Xu, Alexandra Brintrup

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) The Alan Turing Institute(艾伦·图灵研究所)

AI总结 针对合成供应链数据需保持操作逻辑一致性的问题,提出TabKG框架,通过构建列关系知识图谱并利用多LLM集成验证关系,结合潜在扩散模型生成逻辑一致的表格数据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26785 2026-05-27 cs.CL cs.AI

EmoDistill: Offline Emotion Skill Distillation for Language Model Agents in Adversarial Negotiation

EmoDistill: 对抗性谈判中语言模型代理的离线情感技能蒸馏

Yunbo Long, Haolang Zhao, Lukas Beckenbauer, Liming Xu, Alexandra Brintrup

机构 * University of Cambridge(剑桥大学) Technical University of Munich(慕尼黑技术大学) Exiger LLC The Alan Turing Institute(艾伦·图灵研究所)

AI总结 提出EmoDistill离线框架,通过隐式Q学习选择情感和低秩适应策略表达情感,蒸馏情感谈判技能到语言模型代理,在四个高风险谈判领域取得最高效用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06609 2026-05-27 cs.LG cs.CV

Training-Free Vector Quantization via Gaussian VAEs

基于高斯VAE的无训练向量量化

Tongda Xu, Wendi Zheng, Jiajun He, Jose Miguel Hernandez-Lobato, Yan Wang, Ya-Qin Zhang, Jie Tang

机构 * AIR, Tsinghua University(清华空气研究院) CST, Tsinghua University(清华计算机研究所) University of Cambridge(剑桥大学)

AI总结 提出Gaussian Quant (GQ)方法,通过约束训练高斯VAE并直接转换为VQ-VAE,无需额外训练,在UNet和ViT架构上优于现有VQ-VAE。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04310 2026-05-27 cs.AI

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

EvoEmo:面向多轮价格谈判中对抗性LLM智能体的进化情感策略

Yunbo Long, Liming Xu, Lukas Beckenbauer, Yuhan Liu, Alexandra Brintrup

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) Rotman School of Management, University of Toronto(多伦多大学罗特曼管理学院) TUM School of Management, Technical University of Munich(慕尼黑技术大学管理学院) The Alan Turing Institute, London, UK(伦敦阿尔安·图灵研究院)

AI总结 提出EvoEmo进化强化学习框架,通过将情感状态转移建模为马尔可夫决策过程并采用种群遗传优化,动态优化多轮谈判中的情感表达,显著提升LLM智能体的谈判成功率、效率和买家节省。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06843 2026-05-27 cs.CL cs.AI

Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning

自信号驱动的多LLM辩论以实现高效准确的推理

Xuhang Chen, Zhifan Song, Deyi Ji, Shuo Gao, Lanyun Zhu

机构 * University of Cambridge(剑桥大学) Sorbonne Université(索邦大学) University of Science and Technology of China(中国科学技术大学) Beihang University(北航大学) Nanyang Technological University(南洋理工大学)

AI总结 提出一种利用模型级置信度和token级语义焦点两种自信号来自适应引导多LLM辩论过程的方法,在提高准确性的同时减少token消耗。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18728 2026-05-27 cs.LG cs.AI

Message-Passing State-Space Models: Improving Graph Learning with Modern Sequence Modeling

消息传递状态空间模型:利用现代序列建模改进图学习

Andrea Ceni, Alessio Gravina, Claudio Gallicchio, Davide Bacciu, Carola-Bibiane Schonlieb, Moshe Eliasof

机构 * University of Pisa(帕尔米斯大学) University of Cambridge(剑桥大学)

AI总结 提出MP-SSM,将现代状态空间模型的核心计算嵌入消息传递神经网络,实现静态和时序图上的高效、置换等变和长程信息传播,并通过精确敏感性分析刻画深层信息流问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26081 2026-05-26 cs.AI

VeriTrace: Evolving Mental Models for Deep Research Agents

VeriTrace:深度研究智能体的心智模型演化

Haolang Zhao, Yunbo Long, Lukas Beckenbauer, Alexandra Brintrup

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) TUM School of Management, Technical University of Munich(慕尼黑技术大学管理学院) The Alan Turing Institute(艾伦·图灵研究所)

AI总结 针对深度研究智能体面临的信息不确定性,提出VeriTrace认知图框架,通过显式反馈循环(解释更新、偏差反馈、模式修正)来演化心智模型,在DeepResearch Bench和DeepConsult上取得显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.23082 2026-05-26 stat.ML cs.AI cs.LG

KAPLAN: Kolmogorov-Arnold Prognostic Learnable Activation Networks for Survival Analysis

KAPLAN: 用于生存分析的Kolmogorov-Arnold可预测可学习激活网络

Stelios Boulitsakis Logothetis, Angela Wood, Pietro Liò

机构 * University of Cambridge(剑桥大学)

AI总结 提出KAPLAN-HR模型,利用B样条Kolmogorov-Arnold网络非参数估计条件风险函数,通过深层架构自动捕捉交互和时变效应,并证明其收敛速率仅依赖于表示平滑性,从而缓解维度灾难,在六个临床数据集上达到或超越现有方法。

Comments 9 pages, 3 figures, 13 supplementary pages. Submitted to NeurIPS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25446 2026-05-26 cs.AI cs.LG

A Signal-Language Foundation Model for Broad-Spectrum Cardiovascular Assessment from Routine Electrocardiography

面向常规心电图广谱心血管评估的信号-语言基础模型

Ziqing Yu, Yuhui Tao, Jiayu Huo, Lei Pan, Zilong Xiao, Juecheng Chen, Xiao Li, Jianxuan Li, You Zhou, Zhixing Li, Cong Wang, Beijian Zhang, Chen Chen, Hongyang Lu, Konstantinos Patlatzoglou, Daniel B. Kramer, Jonathan W. Waks, Yangang Su, Fu Siong Ng, Shuo Wang, Yixiu Liang, Junbo Ge

机构 * Department of Cardiology, Zhongshan Hospital of Fudan University(复旦大学中山医院心内科) Shanghai Institute of Cardiovascular Diseases, National Clinical Research Centre for Interventional Medicine(上海心血管病研究所,国家介入医学临床研究中心) Digital Medical Research Center, School of Basic Medical Sciences, Fudan University(复旦大学基础医学研究院数字医疗研究中心) Shanghai Key Laboratory of Medical Imaging Computing and Computer Assisted Intervention(上海医学影像计算与计算机辅助手术重点实验室) National Heart and Lung Institute, Imperial College London, Hammersmith Hospital, Du Cane Road(伦敦帝国学院国家心肺研究所,哈马舍姆医院,杜肯路) Department of Cardiology, Shanghai Geriatric Medical Center(上海老年医学中心心内科) Cardiac Rhythm Management, Medtronic Technology Center, Medtronic (Shanghai) Ltd.(美敦力技术中心,美敦力(上海)有限公司,心律管理部) Richard A. and Susan F. Smith Center for Outcomes Research in Cardiology, Beth Israel Deaconess Medical Center, Harvard Medical School(哈佛医学院比尔·德·阿克谢心脏结局研究中心,贝斯以色列·德aconess医疗中心) Harvard-Thorndike Electrophysiology Institute, Beth Israel Deaconess Medical Center, Harvard Medical School(哈佛-托尔恩迪克电生理研究所,贝斯以色列·德aconess医疗中心,哈佛医学院) Department of Cardiology, Imperial College Healthcare NHS Trust(伦敦帝国学院医疗信托心内科部) Department of Cardiology, Chelsea and Westminster NHS Foundation Trust(切尔西和温斯洛医院 NHS 基础信托心内科部) Department of Computer Science and Technology, University of Cambridge(剑桥大学计算机科学与技术系)

AI总结 提出ECGCLIP信号-语言对比学习框架,通过大规模心电图-报告预训练,在89项下游任务中超越基线,实现对常见心律失常、超声心动图靶标及罕见心脏病的广谱评估。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25262 2026-05-26 cs.CV

Semantics-Guided Multimodal Masked Autoencoder Pretraining for 3D BEV Object Detection

语义引导的多模态掩码自编码器预训练用于3D BEV目标检测

Prabuddhi Wariyapperuma, Rajitha de Silva, Marc Hanheide, Thomas Bohné, Leonardo Guevara

机构 * University of Lincoln, Lincoln Centre for Autonomous Systems(林肯大学,林肯自主系统中心) University of Cambridge, Institute for Manufacturing, Department of Engineering(剑桥大学,制造研究所,工程系)

AI总结 提出语义引导的多模态掩码自编码器框架,通过语义引导的LiDAR体素掩码和辅助点语义解码分支,在预训练中注入语义信息,提升3D BEV目标检测性能。

Comments Accepted at the ICRA 2026 Workshop on Semantics for Reliable Robot Autonomy (SRRA) as a lightning talk and poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02979 2026-05-26 cs.CL cs.LG

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning

CPMobius: 无数据强化学习的迭代式教练-玩家推理

Ran Li, Zeyuan Liu, Yinghao Chen, Bingxiang He, Jiarui Yuan, Zixuan Fu, Weize Chen, Jinyi Hu, Chen Qian, Zhiyuan Liu, Maosong Sun

机构 * Tsinghua University(清华大学) University of Cambridge(剑桥大学) Shanghai Jiao Tong University(上海交通大学)

AI总结 提出CPMobius协作式教练-玩家范式,通过无外部数据的合作优化循环提升数学推理能力,在Qwen2.5-Math-7B-Instruct上总体准确率提升4.9%,OOD准确率提升5.4%。

Comments Accepted to the ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24741 2026-05-26 math.ST cs.IT cs.LG math.IT stat.ML stat.TH

On the Sample Complexity of Robust Binary Hypothesis Testing

关于鲁棒二元假设检验的样本复杂度

Shankar Vallinayagam, Ankit Pensia, Varun Jog

机构 * Department of Pure Mathematics and Mathematical Statistics, University of Cambridge(剑桥大学纯数学与数学统计系) Department of Statistics, Carnegie Mellon University(卡内基梅隆大学统计系)

AI总结 研究在三种污染模型下鲁棒二元假设检验的样本复杂度,证明最不利分布的存在性并给出显式公式,揭示样本复杂度对污染参数的不稳定性,并建立不同模型间样本复杂度的可比性。

Comments Comments welcome

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24697 2026-05-26 cs.CL cs.AI

The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models

路径很重要:学习扩散语言模型的令牌提交策略

Bohang Sun, Max Zhu, Francesco Caso, Jindong Gu, Junchi Yu, Philip Torr, Pietro Liò, Jialin Yu

机构 * Department of Computer Science and Technology, University of Cambridge(计算机科学与技术系,剑桥大学) Department of Engineering Science, University of Oxford(工程科学系,牛津大学)

AI总结 本文提出TraceLock,一种轻量级可插拔控制器,通过学习可复用的轨迹状态策略来优化扩散语言模型中的令牌提交决策,从而改善质量与步数之间的权衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10652 2026-05-26 q-bio.QM cs.LG

Querying structural and functional niches on spatial transcriptomics data

查询空间转录组数据中的结构和功能生态位

Mo Chen, Minsheng Hao, Xinquan Liu, Lin Deng, Peng Liu, Chen Li, Dongfang Wang, Kui Hua, Liang Guo, Xuegong Zhang, Lei Wei

机构 * MOE Key Laboratory of Bioinformatics and Bioinformatics Division of BNRIST, Department of Automation, Tsinghua University(生物信息学教育部重点实验室和北京理工大学生物信息学分部,自动化系,清华大学) Center for Synthetic and Systems Biology, School of Life Sciences and School of Medicine, Tsinghua University(合成与系统生物学中心,生命科学学院和医学学院,清华大学) Department of Thoracic Surgery, National Cancer Center/National Clinical Research Center for Cancer/Cancer Hospital, Chinese Academy of Medical Sciences and Peking Union Medical College(胸外科部门,国家癌症中心/国家癌症临床研究中心/癌症医院,中国医学科学院和北京协和医学院) Peking Union Medical College, Chinese Academy of Medical Sciences(北京协和医学院,中国医学科学院) Biomedical Pioneering Innovation Center (BIOPIC), Peking University(生物医学前瞻性创新中心(BIOPIC),北京大学) Cancer Research UK Cambridge Institute, University of Cambridge(英国癌症研究Cambridge研究所,剑桥大学) Department of Immunology, School of Basic Medical Sciences, Harbin Medical University(免疫学部门,基础医学学院,哈尔滨医科大学) Zhongguancun Academy, Beijing, China(中关村学院,北京,中国) Zhongguancun Institute of Artificial Intelligence, Beijing, China(中关村人工智能研究院,北京,中国)

AI总结 提出QueST方法,通过子图建模和对比学习查询空间转录组样本中的相似生态位,有效捕捉异质环境中的生态位结构并跨平台泛化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24524 2026-05-26 cs.LG cs.CL q-bio.NC

What Are We Actually Decoding? Source Attribution for Non-Invasive Brain-to-Language Retrieval

我们究竟在解码什么?非侵入式脑到语言检索的源归因

Xinyu Zhang, Sichao Liu, Runhao Lu, Alexandra Woolgar, Lihui Wang

机构 * KTH(瑞典皇家理工学院) University of Cambridge(剑桥大学) EPFL(苏黎世联邦理工学院) Karolinska Institutet(Karolinska研究所) McGill University(麦吉尔大学)

AI总结 针对非侵入式神经语言解码中结果被非刺激诱发源(如解码器先验、嵌入度量、信号时长等)膨胀的问题,提出一个审计框架,通过结构捷径、窗口级刺激锁定证据和跨窗口上下文聚合三种源分离,并引入组上下文偏差(GCB)作为可控的源归因干预,实现性能的源归因而非仅报告。

Comments 35 pages, 7 figures, 25 tables

详情

展开后加载摘要…

URL PDF HTML 收藏