arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

共收录 2801
2602.07205 2026-02-10 cs.LG cs.GT stat.ML

Online Learning for Uninformed Markov Games: Empirical Nash-Value Regret and Non-Stationarity Adaptation

在线学习无信息马尔可夫游戏:经验纳什价值遗憾与非平稳适应

Junyan Liu, Haipeng Luo, Zihan Zhang, Lillian J. Ratliff

机构 * University of Washington(华盛顿大学) University of Southern California(南加州大学) Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 本文提出了一种参数无关算法,通过经验纳什价值遗憾度量,在对手非平稳性适应下实现O(min{√K + (CK)^{1/3},√LK})的遗憾界。

Comments 36 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06949 2026-02-09 cs.RO cs.AI cs.CV cs.LG

DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos

DreamDojo:从大规模人类视频中学习通用机器人世界模型

Shenyuan Gao, William Liang, Kaiyuan Zheng, Ayaan Malik, Seonghyeon Ye, Sihyun Yu, Wei-Cheng Tseng, Yuzhu Dong, Kaichun Mo, Chen-Hsuan Lin, Qianli Ma, Seungjun Nah, Loic Magne, Jiannan Xiang, Yuqi Xie, Ruijie Zheng, Dantong Niu, You Liang Tan, K. R. Zentner, George Kurian, Suneel Indupuru, Pooya Jannaty, Jinwei Gu, Jun Zhang, Jitendra Malik, Pieter Abbeel, Ming-Yu Liu, Yuke Zhu, Joel Jang, Linxi "Jim" Fan

机构 * NVIDIA HKUST(香港科技大学) UC Berkeley(加州大学伯克利分校) Stanford(斯坦福大学) KAIST(韩国科学技术院) UofT(多伦多大学) UCSD(加州大学圣地亚哥分校) UT Austin(德克萨斯大学奥斯汀分校)

AI总结 DreamDojo通过大规模人类视频学习通用机器人世界模型,解决动作标签稀缺问题,实现高精度物理理解和实时交互。

Comments Project page: https://dreamdojo-world.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06674 2026-02-09 cs.CV cs.HC cs.LG

CytoCrowd: A Multi-Annotator Benchmark Dataset for Cytology Image Analysis

CytoCrowd:一种用于细胞学图像分析的多标注基准数据集

Yonghao Si, Xingyuan Zeng, Zhao Chen, Libin Zheng, Caleb Chen Cao, Lei Chen, Jian Yin

机构 * Sun Yat-sen University(中山大学) Hong Kong University of Science and Technology(香港科技大学)

AI总结 CytoCrowd是一个包含446张高分辨率细胞学图像的数据集,提供四个病理学家的冲突注释和一个资深专家的金牌标准,用于评估注释聚合算法和计算机视觉任务的基准测试。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06453 2026-02-09 cs.LG

On the Plasticity and Stability for Post-Training Large Language Models

关于后训练大语言模型的可塑性与稳定性

Wenwen Qiang, Ziyin Gu, Jiahuan Zhou, Jie Hu, Jingyao Wang, Changwen Zheng, Hui Xiong

机构 * Institute of Software Chinese Academy of Sciences, Beijing, China(中国科学院软件研究所) University of the Chinese Academy of Sciences, Beijing, China(中国科学院大学) Wangxuan Institute of Computer Technology, Peking University, Beijing, China(北京大学王轩计算机技术研究所) Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou), China(香港科技大学(广州)人工智能研究所) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology Hong Kong SAR, China(香港科技大学(香港特别行政区)计算机科学与工程系)

AI总结 本文提出PCR框架,通过概率方法解决GRPO中可塑性与稳定性之间的几何冲突,提升训练稳定性与推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06319 2026-02-09 cs.AI

Exposing Weaknesses of Large Reasoning Models through Graph Algorithm Problems

通过图算法问题揭示大推理模型的弱点

Qifan Zhang, Jianhao Ruan, Aochuan Chen, Kang Zeng, Nuo Chen, Jing Tang, Jia Li

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

AI总结 GrAlgoBench通过图算法问题揭示大推理模型在长上下文推理和过度思考方面的不足,为改进推理研究提供严格测试平台。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05885 2026-02-09 cs.LG cs.AI cs.CL

Dr. Kernel: Reinforcement Learning Done Right for Triton Kernel Generations

Dr. Kernel:为Triton内核生成正确强化学习

Wei Liu, Jiawei Xu, Yingru Li, Longtao Zheng, Tianjian Li, Qian Liu, Junxian He

机构 * HKUST(香港科技大学) TikTok CUHK(SZ)(香港中文大学(深圳)) NTU(国立科技大学)

AI总结 Dr. Kernel通过强化学习方法优化内核生成,其模型在Kernelbench测试中达到与Claude-4.5-Sonnet相当的性能,并在速度提升方面超越其他模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05444 2026-02-09 cs.CL

Causal Front-Door Adjustment for Robust Jailbreak Attacks on LLMs

因果前门调整用于对抗大语言模型的鲁棒性劫持攻击

Yao Zhou, Zeen Song, Wenwen Qiang, Fengge Wu, Shuyi Zhou, Changwen Zheng, Hui Xiong

机构 * Institute of Software Chinese Academy of Sciences, Beijing, China(中国科学院软件研究所) University of Chinese Academy of Sciences, Beijing, China(中国科学院大学) Institute of Information Engineering Chinese Academy of Sciences, Beijing, China(中国科学院信息工程研究所) Hong Kong University of Science and Technology, China, Hong Kong, China(香港科学与技术大学)

AI总结 本文提出CFA²攻击方法,通过因果前门准则和稀疏自编码器实现对大语言模型的鲁棒性劫持,提升攻击成功率并提供机制解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00169 2026-02-09 cond-mat.mtrl-sci cs.AI

Towards Agentic Intelligence for Materials Science

迈向材料科学的代理智能

Huan Zhang, Yizhan Li, Wenhao Huang, Ziyu Hou, Yu Song, Xuye Liu, Farshid Effaty, Jinya Jiang, Sifan Wu, Qianggang Ding, Izumi Takahara, Leonard R. MacGillivray, Teruyasu Mizoguchi, Tianshu Yu, Lizi Liao, Yuyu Luo, Yu Rong, Jia Li, Ying Diao, Heng Ji, Bang Liu

机构 * DIRO & Institut Courtois, Université de Montréal(蒙特利尔大学DIRO与Courtois研究所) Mila – Quebec AI Institute(魁北克AI研究所) University of Waterloo(滑铁卢大学) Université de Sherbrooke(Sherbrooke大学) University of California, San Diego(加州大学圣地亚哥分校) The University of Tokyo, Institute of Industrial Science(东京大学工业科学研究所) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Singapore Management University(新加坡管理学院) The Hong Kong University of Science and Technology, Guangzhou(香港科学与技术大学(广州)) Alibaba DAMO Academy(阿里巴巴达摩院) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) CIFAR AI Chair(CIFAR人工智能主席)

AI总结 本文提出以流程为中心的代理系统框架,通过整合人工智能与材料科学,推动材料发现的自主化与智能化。

Comments 81 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02075 2026-02-09 cs.CE cs.LG

MDAgent2: Large Language Model for Code Generation and Knowledge Q&A in Molecular Dynamics

MDAgent2:用于分子动力学中代码生成和知识问答的大型语言模型

Zhuofan Shi, Hubao A, Yufei Shao, Dongliang Huang, Hongxu An, Chunxiao Xin, Haiyang Shen, Zhenyu Wang, Yunshan Na, Gang Huang, Xiang Jing

机构 * Peking University(北京大学) National Key Laboratory of Data Space Technology and System(国家数据空间技术与系统重点实验室) The Hong Kong University of Science and Technology(香港科学与技术大学) Liaoning Technical University(辽宁技术大学) Wenjing Future Lab (Beijing) Technology Co., Ltd(文景未来实验室(北京)科技有限公司)

AI总结 MDAgent2是首个能同时进行分子动力学知识问答和代码生成的端到端框架,通过领域特定数据集和强化学习方法提升性能。

Comments 24 pages,4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03054 2026-02-09 cs.LG cs.AI

Calibration and Transformation-Free Weight-Only LLMs Quantization via Dynamic Grouping

无需校准和转换的权重-only LLMs 量化 via 动态分组

Xinzhe Zheng, Zhen-Qun Yang, Zishan Liu, Haoran Xie, S. Joe Qin, Arlene Chen, Fangzhen Lin

机构 * Division of Artificial Intelligence, School of Data Science, Lingnan University, Hong Kong, China(岭南大学人工智能学院) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology, Hong Kong, China(香港理工大学计算机科学与工程系) Department of Computing, The Hong Kong Polytechnic University, Hong Kong, China(香港理工大学计算机系) Xiaoi Robot Inc., Shanghai, China(小蚁机器人有限公司)

AI总结 MSB提出一种无需校准和转换的低比特PTQ方法,通过动态分组优化实现多尺度量化,提升LLMs在内存和计算约束下的性能。

Comments 34 pages, 10 figures. Version 3 corrects the bit-length error and adds new experiments and analysis; the core methodology remains unchanged. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11781 2026-02-09 cs.LG

Multi-Order Wavelet Derivative Transform for Deep Time Series Forecasting

多阶小波导数变换用于深度时间序列预测

Ziyu Zhou, Jiaxi Hu, Qingsong Wen, James T. Kwok, Yuxuan Liang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Squirrel AI Learning

AI总结 多阶小波导数变换通过提取时间感知模式,提升深度时间序列预测的精度与效率。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04010 2026-02-09 cs.LG cs.AI cs.CL cs.NE

Hyperbolic Fine-Tuning for Large Language Models

双曲微调用于大语言模型

Menglin Yang, Ram Samarth B B, Aosong Feng, Bo Xiong, Jihong Liu, Irwin King, Rex Ying

机构 * HKUST(GZ)(香港科技大学(广州)) HKUST(香港科技大学) Indian Institute of Science(印度科学研究院) Yale University(耶鲁大学) Stanford University(斯坦福大学) The Chinese University of Hong Kong(香港中文大学)

AI总结 HypLoRA通过在双曲空间中进行低秩适应,提升大语言模型在算术和常识推理任务中的性能。

Comments NeurIPS 2025; https://github.com/marlin-codes/HypLoRA

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14069 2026-02-09 cs.CV

Self-Supervised Video Representation Learning in a Heuristic Decoupled Perspective

基于启发式解耦视角的自监督视频表示学习

Zeen Song, Wenwen Qiang, Changwen Zheng, Hui Xiong, Gang Hua

机构 * National Key Laboratory of Space Integrated Information System, Institute of Software Chinese Academy of Sciences(空间信息集成国家重点实验室,软件研究所中国科学院) University of Chinese Academy of Sciences(中国科学院大学) Hong Kong University of Science and Technology(香港理工大学) Dolby Laboratories Inc(杜比实验室有限公司) Xi’an Jiaotong University(西安交通大学)

AI总结 本文提出BOD-VCL方法,通过解耦静态和动态语义,提升视频对比学习的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05798 2026-02-06 stat.ME cs.LG eess.SP stat.ML

Learning False Discovery Rate Control via Model-Based Neural Networks

通过基于模型的神经网络学习虚假发现率控制

Arnau Vilella, Jasin Machkour, Michael Muma, Daniel P. Palomar

机构 * The Hong Kong University of Science and Technology(香港理工大学) Technische Universität Darmstadt(德累斯顿理工大学)

AI总结 本文提出基于模型的神经网络方法,改进T-Rex选择器框架,以更精确地控制FDR并提高发现能力。

Comments Accepted to IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05479 2026-02-06 cs.AI

Phi-Former: A Pairwise Hierarchical Approach for Compound-Protein Interactions Prediction

Phi-Former:一种用于化合物-蛋白质相互作用预测的成对分层方法

Zhe Wang, Zijing Liu, Chencheng Xu, Yuan Yao

机构 * Hong Kong University of Science and Technology(香港科技大学) International Digital Economy Academy (IDEA)(国际数字经济学院(IDEA)) Princeton University(普林斯顿大学)

AI总结 Phi-Former通过成对分层方法提升化合物-蛋白质相互作用预测的准确性与可解释性。

Comments Accepted to BIBM 2025. 6 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05393 2026-02-06 cs.CL cs.LG

Late-to-Early Training: LET LLMs Learn Earlier, So Faster and Better

后期到早期训练:LET LLMs 早期学习,从而更快且更好

Ji Zhao, Yufei Gu, Shitong Shao, Xun Zhou, Liang Xiang, Zeke Xie

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

AI总结 本文提出LET范式,通过利用已有小型预训练模型加速大模型训练,实现更快的训练速度和更高的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07253 2026-02-06 cs.LG cs.CV

Alignment of Diffusion Models: Fundamentals, Challenges, and Future

扩散模型对齐:基础、挑战与未来

Buhua Liu, Shitong Shao, Bao Li, Lichen Bai, Zhiqiang Xu, Haoyi Xiong, James Kwok, Sumi Helal, Zeke Xie

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) Baidu Inc.(百度公司) The University of Bologna(博洛尼亚大学)

AI总结 本文综述了扩散模型对齐的基础、挑战及未来方向,探讨了对齐技术、评估方法及当前挑战的解决方案。

Comments Accepted at ACM Computing Surveys. 35 pages, 5 figures, 4 tables. Paper List: github.com/xie-lab-ml/awesome-alignment-of-diffusion-models

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05323 2026-02-06 cs.LG cs.AI

GAS: Enhancing Reward-Cost Balance of Generative Model-assisted Offline Safe RL

GAS: 提升生成模型辅助的离线安全强化学习的奖励-成本平衡

Zifan Liu, Xinran Li, Shibo Chen, Jun Zhang

机构 * The Hong Kong University of Science and Technology, Hong Kong SAR, China(香港科学与技术大学) South China University of Technology, Guangdong, China(华南理工大学)

AI总结 GAS通过增强数据集和引入目标函数,提升生成模型辅助离线安全强化学习中奖励与成本的平衡能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04926 2026-02-06 cs.DB cs.CL cs.LG

Pruning Minimal Reasoning Graphs for Efficient Retrieval-Augmented Generation

对检索增强生成进行最小推理图剪枝以提高效率

Ning Wang, Kuanyan Zhu, Daniel Yuehwoon Yee, Yitang Gao, Shiying Huang, Zirun Xu, Sainyam Galhotra

机构 * Cornell University(康奈尔大学) University of Cambridge(剑桥大学) The University of Hong Kong(香港大学) HKUST(香港科技大学) University of British Columbia(不列颠哥伦比亚大学)

AI总结 AutoPrunedRetriever通过最小推理图剪枝提升检索增强生成效率,实现更高效的知识密集型任务处理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19243 2026-02-06 cs.CV

VisionDirector: Vision-Language Guided Closed-Loop Refinement for Generative Image Synthesis

VisionDirector: 生成图像合成中的视觉-语言引导闭环细化

Meng Chu, Senqiao Yang, Haoxuan Che, Suiyun Zhang, Xichen Zhang, Shaozuo Yu, Haokun Gui, Zhefan Rao, Dandan Tu, Rui Liu, Jiaya Jia

机构 * The Hong Kong University of Science and Technology(香港理工大学) The Chinese University of Hong Kong(香港中文大学) Huawei Research(华为研究)

AI总结 VisionDirector通过视觉-语言引导闭环细化方法,提升生成图像合成中多目标任务的完成度和质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09465 2026-02-05 cs.AI

EvoFSM: Controllable Self-Evolution for Deep Research with Finite State Machines

EvoFSM: 通过有限状态机实现可控的自我进化用于深度研究

Shuo Zhang, Chaofa Yuan, Ryan Guo, Xiaomin Yu, Rui Xu, Zhangquan Chen, Zinuo Li, Zhi Yang, Shuhao Guan, Zhenheng Tang, Sen Hu, Liwen Zhang, Ronghao Chen, Huacan Wang

机构 * QuantaAlpha HKUST(GZ)(香港科技大学(广州)) FDU(福建大学) THU(清华大学) SUFE(上海财经大学) UCD(都柏林大学) HKUST(香港科技大学) PKU(北京大学) UCAS(中国科学技术大学)

AI总结 EvoFSM通过有限状态机实现可控自我进化,提升智能体在开放性问题中的适应性与稳定性,有效解决传统方法的不稳定性与幻觉问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09142 2026-02-05 cs.LG cs.CL

EvasionBench: A Large-Scale Benchmark for Detecting Managerial Evasion in Earnings Call Q&A

EvasionBench: 一个大规模基准用于检测盈余电话会议问答中的管理层回避

Shijian Ma, Yan Lin, Yi Yang

机构 * Hong Kong University of Science and Technology(香港科学与技术大学) University of Macau(澳门大学)

AI总结 EvasionBench通过多模型共识框架和40亿参数分类器,首次大规模评估管理层在盈余电话会议中回避性回应的检测能力。

Comments Major revision. Title and abstract updated to better reflect the refined results. Shijian Ma and Yan Lin contributed equally. Corresponding author: Yan Lin; Project page: https://iiiiqiiii.github.io/EvasionBench/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.16793 2026-02-05 cs.RO

PhysBrain: Human Egocentric Data as a Bridge from Vision Language Models to Physical Intelligence

PhysBrain: 人眼视角数据作为视觉语言模型到物理智能的桥梁

Xiaopeng Lin, Shijie Lian, Bin Yu, Ruoqi Yang, Zhaolong Shen, Changti Wu, Yuzhuo Miao, Yurun Jin, Yukun Shi, Jiyan He, Cong Huang, Bojun Cheng, Kai Chen

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Zhongguancun Academy(中关村学院) Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院) Harbin Institute of Technology(哈尔滨工业大学) Huazhong University of Science and Technology(华中科技大学)

AI总结 PhysBrain通过将人类眼动视频转化为多级具身监督,提升机器人在眼动感知和长期规划中的能力,实现从人类视角到机器人控制的有效迁移。

Comments 21 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20322 2026-02-05 cs.CV

Dynamic Pyramid Network for Efficient Multimodal Large Language Model

动态金字塔网络用于高效多模态大语言模型

Hao Ai, Kunyi Wang, Zezhou Wang, Hao Lu, Jin Tian, Yaxin Luo, Peng Xing, Jen-Yuan Huang, Huaxia Li, Gen luo

机构 * Beihang University(北航大学) Shanghai AI Laboratory(上海人工智能实验室) KAUST(卡塔尔大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Technical University Of Denmark(丹麦技术大学) Nanjing University of Science and Technology(南京理工大学) Peking University(北京大学) Xiaohongshu Inc(小红书公司)

AI总结 动态金字塔网络通过分层结构和动态池化专家提升多模态大语言模型的效率与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11733 2026-02-05 cs.CY cs.AI cs.CL cs.HC

LLM Agents for Education: Advances and Applications

教育中的LLM代理:进展与应用

Zhendong Chu, Shen Wang, Jian Xie, Tinghui Zhu, Yibo Yan, Jinheng Ye, Aoxiao Zhong, Xuming Hu, Jing Liang, Philip S. Yu, Qingsong Wen

机构 * Squirrel Ai Learning Fudan University(复旦大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Tsinghua University(清华大学) University of Illinois Chicago(伊利诺伊大学芝加哥分校)

AI总结 本文综述了LLM代理在教育中的应用进展,探讨了其技术实现、挑战及在不同教育领域的应用。

Comments Accepted by EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04240 2026-02-05 cs.CV cs.LG cs.RO

SPOT-Occ: Sparse Prototype-guided Transformer for Camera-based 3D Occupancy Prediction

SPOT-Occ: 基于稀疏原型的Transformer用于基于相机的3D占用预测

Suzeyu Chen, Leheng Li, Ying-Cong Chen

机构 * Artificial Intelligence Thrust, The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学人工智能方向(广州)) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学计算机科学与工程系)

AI总结 SPOT-Occ提出了一种基于稀疏原型的Transformer解码器,通过引导特征选择和聚焦聚合提升基于相机的3D占用预测的效率与精度。

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04214 2026-02-05 cs.RO

ALORE: Autonomous Large-Object Rearrangement with a Legged Manipulator

ALORE:具有腿部机械臂的自主大物体重排

Zhihai Bi, Yushan Zhang, Kai Chen, Guoyang Zhao, Yulin Li, Jun Ma

机构 * Robotics and Autonomous Systems Thrust, The Hong Kong University of Science and Technology (Guangzhou)(机器人与自主系统方向,香港科技大学(广州))

AI总结 ALORE通过分层强化学习和任务规划框架,实现了腿部机械臂在复杂环境中的多物体自主重排,展示了高效稳定的操作性能和广泛的应用潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21005 2026-02-05 stat.ML cs.LG math.PR

Learning from Neighbors with PHIBP: Predicting Infectious Disease Dynamics in Data-Sparse Environments

通过PHIBP学习邻居:在数据稀疏环境中预测传染病动态

Edwin Fong, Lancelot F. James, Juho Lee

机构 * Department of Statistics and Actuarial Science, HKU(香港大学统计与精算科学系) Department of ISOM, HKUST(香港科技大学工业系统与管理系) The Graduate School of AI, KAIST(韩国科学技术院人工智能研究生院)

AI总结 PHIBP通过系统地借鉴相关地区统计强度,有效处理稀疏计数数据,提升传染病预测的准确性和流行病学洞察。

Comments v2: Revised version incorporating peer review feedback from book chapter submission. Clarifies modeling objectives for infectious disease prediction and situates the work within a three-paper PHIBP framework, highlighting suitability for future AI/LLM plug-and-play model specification

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03815 2026-02-04 cs.CV cs.LG

Fast-Slow Efficient Training for Multimodal Large Language Models via Visual Token Pruning

通过视觉标记修剪实现多模态大语言模型的快慢高效训练

Dingkun Zhang, Shuhan Qi, Yulin Wu, Xinyu Xiao, Xuan Wang, Long Chen

机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Hong Kong University of Science and Technology(香港科技大学)

AI总结 DualSpeed通过快慢双模式实现多模态大语言模型的高效训练,提升训练速度并保持性能

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18939 2026-02-04 cs.LG cs.AI

Damba-ST: Domain-Adaptive Mamba for Efficient Urban Spatio-Temporal Prediction

Damba-ST: 域自适应Mamba用于高效的都市时空预测

Rui An, Yifeng Zhang, Ziran Liang, Wenqi Fan, Yuxuan Liang, Xuequn Shang, Qing Li

机构 * Northwestern Polytechnical University(西北工业大学) The Hong Kong University of Science and Technology (Guangzhou)(香港理工大学(广州))

AI总结 Damba-ST通过域自适应Mamba模型提升都市时空预测的效率与泛化能力,解决跨域泛化难题。

Comments Accepted by ICDE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏