arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

共收录 2801
2602.03400 2026-02-04 cs.SE cs.AI

Precision in Practice: Knowledge Guided Code Summarizing Grounded in Industrial Expectations

实践中的精确性:基于工业期望的知识引导代码摘要

Jintai Li, Songqiang Chen, Shuo Jin, Xiaoyuan Xie

机构 * School of Computer Science, Wuhan University, China(武汉大学计算机科学学院) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology, China(香港科学与技术大学计算机科学与工程系) Department of Computer Science(计算机科学系)

AI总结 ExpSum 通过整合元数据抽象和领域知识检索,生成符合工业期望的代码摘要,显著提升摘要质量与实用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03237 2026-02-04 cs.LG cs.CL

Merging Beyond: Streaming LLM Updates via Activation-Guided Rotations

超越合并:通过激活引导的旋转进行流式LLM更新

Yuxuan Yao, Haonan Sheng, Qingsong Lv, Han Wu, Shuqi Liu, Zehua Liu, Zengyan Liu, Jiahui Gao, Haochen Tan, Xiaojin Fu, Haoli Bai, Hing Cheung So, Zhijiang Guo, Linqi Song

机构 * City University of Hong Kong, Hong Kong SAR(香港城市大学) Tsinghua University(清华大学) Huawei Noah’s Ark Lab, Hong Kong SAR(华为诺亚实验室(香港)) University of Hong Kong(香港大学) Hong Kong University of Science and Technology (Guangzhou)(香港理工大学(广州))

AI总结 本文提出ARM策略,通过激活引导的旋转实现流式LLM更新,有效超越收敛模型,提供高效适应框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03086 2026-02-04 cs.LG cs.CV

Neural Predictor-Corrector: Solving Homotopy Problems with Reinforcement Learning

神经预测-校正器:利用强化学习解决同伦问题

Jiayao Mai, Bangyan Liao, Zhenjun Zhao, Yingping Zeng, Haoang Li, Javier Civera, Tailin Wu, Yi Zhou, Peidong Liu

机构 * Hunan University(湖南大学) Westlake University(西湖大学) University of Zaragoza(阿拉贡大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 本文提出神经预测-校正器,通过强化学习统一解决同伦问题,实现高效稳定的方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03026 2026-02-04 cs.AI cs.MA

Visual Reasoning over Time Series via Multi-Agent System

通过多智能体系统进行时间序列的视觉推理

Weilin Ruan, Yuxuan Liang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 MAS4TS通过多智能体系统实现时间序列的视觉推理,利用分析-推理-执行范式,结合视觉语言模型和工具链,提升时间序列任务的性能和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02539 2026-02-04 cs.LG cs.CV

How Much Information Can a Vision Token Hold? A Scaling Law for Recognition Limits in VLMs

视觉标记能承载多少信息?VLMs中识别限制的缩放定律

Shuxin Zhuang, Zi Liang, Runsheng Yu, Hongzong Li, Rong Feng, Shiqin Tang, Youzhi Zhang

机构 * City University of Hong Kong(香港城市大学) The Hong Kong Polytechnic University(香港理工大学) The Hong Kong University of Science and Technology(香港理工大学) Centre for Artificial Intelligence and Robotics, Chinese Academy of Sciences(中国科学院人工智能与机器人研究院)

AI总结 本研究通过压力测试揭示了视觉标记的信息上限,提出了一种统一的缩放定律,为优化视觉上下文压缩的效率-精度权衡提供了实证依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02230 2026-02-04 cs.LG cs.AI

SEDformer: Event-Synchronous Spiking Transformers for Irregular Telemetry Time Series Forecasting

SEDformer: 事件同步脉冲变换器用于不规则遥测时间序列预测

Ziyu Zhou, Yuchen Fang, Weilin Ruan, Shiyu Wang, James Kwok, Yuxuan Liang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) University of Electronic Science and Technology of China(电子科学与技术大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 SEDformer是一种基于脉冲变换器的遥测不规则时间序列预测模型,通过事件同步脉冲编码和事件保留下采样模块,提升预测精度并降低能耗。

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02033 2026-02-04 cs.CV cs.AI cs.MM

One Size, Many Fits: Aligning Diverse Group-Wise Click Preferences in Large-Scale Advertising Image Generation

一个尺寸,多种适配:在大规模广告图像生成中对多样化群体点击偏好进行对齐

Shuo Lu, Haohan Wang, Wei Feng, Weizhen Wang, Shen Zhang, Yaoyu Li, Ao Ma, Zheng Zhang, Jingjing Lv, Junjie Shen, Ching Law, Bing Zhan, Yuan Xu, Huizai Yao, Yongcan Yu, Chenyang Si, Jian Liang

机构 * NLPR & MAIS, CASIA(中国科学院长春光学精密机械与物理研究所 & 中国科学院自动化所) School of AI, UCAS(中国科学院大学人工智能学院) HKUST(gz)(香港科技大学) PRLab, NJU(南京大学PRLab)

AI总结 本文提出OSMF框架,通过自适应分组和群组感知多模态模型,解决广告图像生成中用户群体点击偏好多样性的优化问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01077 2026-02-04 cs.CV

PISA: Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers

PISA:分块稀疏注意力使扩散变换器更高效

Haopeng Li, Shitong Shao, Wenliang Zhong, Zikai Zhou, Lichen Bai, Hui Xiong, Zeke Xie

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

AI总结 PISA通过分块稀疏注意力在保持质量的同时显著提升扩散变换器的效率。

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12203 2026-02-04 cs.CV

LazyDrag: Enabling Stable Drag-Based Editing on Multi-Modal Diffusion Transformers via Explicit Correspondence

LazyDrag: 通过显式对应关系在多模态扩散变换器上实现稳定的拖拽编辑

Zixin Yin, Xili Dai, Duomin Wang, Xianfang Zeng, Lionel M. Ni, Gang Yu, Heung-Yeung Shum

机构 * The Hong Kong University of Science and Technology(香港科技大学) StepFun The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 LazyDrag通过显式对应关系实现多模态扩散变换器的稳定拖拽编辑,消除了隐式点匹配依赖,提升生成能力与精确控制。

Comments https://zxyin.github.io/LazyDrag

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09131 2026-02-04 cs.GR cs.AI cs.CV

Training-Free Text-Guided Color Editing with Multi-Modal Diffusion Transformer

无需训练的文本引导颜色编辑与多模态扩散变换器

Zixin Yin, Xili Dai, Ling-Hao Chen, Deyu Zhou, Jianan Wang, Duomin Wang, Gang Yu, Lionel M. Ni, Lei Zhang, Heung-Yeung Shum

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) International Digital Economy Academy(国际数字经济学院) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Tsinghua University(清华大学) Astribot StepFun

AI总结 ColorCtrl通过多模态扩散变换器实现无需训练的文本引导颜色编辑,精准控制颜色属性并保持一致性,优于现有方法和商业模型。

Comments https://zxyin.github.io/ColorCtrl

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12968 2026-02-04 cs.CV cs.RO

OptiPMB: Enhancing 3D Multi-Object Tracking with Optimized Poisson Multi-Bernoulli Filtering

OptiPMB:通过优化的泊松多伯努利滤波增强3D多目标跟踪

Guanhua Ding, Yuxuan Xia, Runwei Guan, Qinchen Wu, Tao Huang, Weiping Ding, Jinping Sun, Guoqiang Mao

机构 * School of Electronic Information Engineering, Beihang University(电子信息工程学院,北京航空航天大学) School of Automation and Intelligent Sensing, Shanghai Jiaotong University(自动化与智能感知学院,上海交通大学) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(计算机科学与工程系,香港科技大学) College of Science and Engineering, James Cook University(科学与工程学院,詹姆斯库克大学) School of Artificial Intelligence and Computer Science, Nantong University(人工智能与计算机科学学院,南通大学) Faculty of Data Science, City University of Macau(数据科学学院,澳门城市大学) Research Laboratory of Smart Driving and Intelligent Transportation Systems, Southeast University(智能驾驶与智能交通系统研究实验室,东南大学)

AI总结 OptiPMB通过优化的泊松多伯努利滤波器和创新设计提升3D多目标跟踪的准确性与性能

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02159 2026-02-03 cs.CL

Focus-dLLM: Accelerating Long-Context Diffusion LLM Inference via Confidence-Guided Context Focusing

Focus-dLLM: 通过置信度引导的上下文聚焦加速长上下文扩散语言模型推理

Lingkun Long, Yushi Huang, Shihao Bai, Ruihao Gong, Jun Zhang, Ao Zhou, Jianlei Yang

机构 * Beihang University(北航) Hong Kong University of Science and Technology(香港理工大学) SenseTime Research(商汤科技研究院)

AI总结 Focus-dLLM通过置信度引导的上下文聚焦技术,实现了长上下文扩散语言模型推理的高效加速。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02156 2026-02-03 cs.CV

LoopViT: Scaling Visual ARC with Looped Transformers

LoopViT: 通过循环变换器扩展视觉ARC

Wen-Jie Shu, Xuerui Qiu, Rui-Jie Zhu, Harold Haodong Chen, Yexin Liu, Harry Yang

机构 * HKUST(香港科技大学) CASIA(中国科学院自动化研究所) UC Santa Cruz(加州大学圣克ruz分校)

AI总结 LoopViT通过循环变换器实现视觉推理的高效扩展,采用权重绑定递归结构和动态退出机制,提升ARC-AGI基准测试性能。

Comments 8 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01965 2026-02-03 cs.CL cs.AI

Breaking the Static Graph: Context-Aware Traversal for Robust Retrieval-Augmented Generation

打破静态图:面向鲁棒检索增强生成的上下文感知遍历

Kwun Hang Lau, Fangyuan Zhang, Boyu Ruan, Yingli Zhou, Qintian Guo, Ruiyuan Zhang, Xiaofang Zhou

机构 * Huawei Hong Kong Research Center(华为香港研究中心) The Hong Kong University of Science and Technology(香港科学与技术大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

AI总结 CatRAG通过上下文感知遍历框架,改进检索增强生成模型,提升多跳查询的推理完整性和证据链恢复能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01877 2026-02-03 cs.LG math.OC

Autocorrelated Optimize-via-Estimate: Predict-then-Optimize versus Finite-sample Optimal

自相关优化-通过估计:预测-然后优化与有限样本最优

Zichun Wang, Gar Goei Loke, Ruiting Zuo

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Durham University Business School(杜伦大学商学院)

AI总结 本文提出了一种自相关优化-通过估计模型,用于在有限样本情况下优化资产组合,表现出优于传统方法的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01825 2026-02-03 stat.ME cs.LG math.OC stat.ML

Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes

通过群鲁棒马尔可夫决策过程从多个来源学习顺序决策

Mingyuan Xu, Zongqi Xia, Tianxi Cai, Doudou Zhou, Nian Si

机构 * Department of Statistics and Data Science, National University of Singapore(新加坡国立大学统计与数据科学系) Department of Neurology, University of Pittsburgh(匹兹堡大学神经病学系) Department of Biostatistics, Harvard T.H. Chan School of Public Health(哈佛大学T.H. Chan公共卫生学院生物统计学系) Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系) Department of Industrial Engineering and Decision Analytics, Hong Kong University of Science and Technology(香港科学与技术大学工业工程与决策分析系)

AI总结 本文提出了一种群鲁棒马尔可夫决策过程框架,通过特征层面的不确定性集和悲观价值迭代算法,从多地点异质数据中学习鲁棒的顺序决策策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10942 2026-02-03 cs.CV

VL-JEPA: Joint Embedding Predictive Architecture for Vision-language

VL-JEPA:面向视觉-语言的联合嵌入预测架构

Delong Chen, Mustafa Shukor, Theo Moutakanni, Willy Chung, Jade Yu, Tejaswi Kasarla, Yejin Bang, Allen Bolourchi, Yann LeCun, Pascale Fung

机构 * Meta FAIR HKUST(香港科技大学) Sorbonne Université(索邦大学) NYU(纽约大学)

AI总结 VL-JEPA通过联合嵌入预测架构,在减少参数的情况下实现更强的视觉-语言性能,支持多种任务且无需架构修改。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13745 2026-02-03 cs.CV

UniCalli: A Unified Diffusion Framework for Column-Level Generation and Recognition of Chinese Calligraphy

UniCalli: 一种统一的扩散框架,用于中文书法的列级生成与识别

Tianshuo Xu, Kai Wang, Zhifei Chen, Leyi Wu, Tianshui Wen, Fei Chao, Ying-Cong Chen

机构 * HKUST(GZ)(香港科技大学(广州)) China University of Geoscience Beijing(中国地质大学(北京)) Xiamen University(厦门大学) HKUST(香港科技大学)

AI总结 UniCalli提出一种统一的扩散框架,通过联合训练实现中文书法列级生成与识别,提升生成质量和识别性能,同时扩展至其他古代文字。

Comments Page: https://envision-research.github.io/UniCalli/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12094 2026-02-03 cs.LG cs.GT

Is This Predictor More Informative than Another? A Decision-Theoretical Comparison

这个预测器比另一个更有信息量吗?一种决策理论的比较

Yiding Feng, Liuhan Qian, Wei Tang

机构 * Hong Kong University of Science and Technology(香港科学与技术大学) The Chinese University of Hong Kong(香港中文大学)

AI总结 本文提出信息量差距的概念,用于比较预测器的决策相关性,提供了一种评估预测模型在不同决策任务中表现的理论框架和实验验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18561 2026-02-03 cs.CV

CoT-RVS: Zero-Shot Chain-of-Thought Reasoning Segmentation for Videos

CoT-RVS:零样本链式推理视频对象分割

Shiu-hong Kao, Yu-Wing Tai, Chi-Keung Tang

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) National University of Singapore(新加坡国立大学) Dartmouth College(达特茅斯学院)

AI总结 CoT-RVS通过零样本链式推理能力,实现了对视频对象的高效分割,无需训练即可处理复杂查询和在线视频流。

Comments Accepted to ICLR 2026. Project page: https://danielshkao.github.io/cot-rvs.html. Code: https://github.com/DanielSHKao/CoT-RVS

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15386 2026-02-03 cs.CL cs.AI

RePPL: Recalibrating Perplexity by Uncertainty in Semantic Propagation and Language Generation for Explainable QA Hallucination Detection

RePPL:通过语义传播和语言生成中的不确定性重新校准困惑度以检测可解释问答幻觉

Yiming Huang, Junyan Zhang, Zihao Wang, Biquan Bie, Yunzhong Qiu, Xuming Hu, Yi R. Fung, Xinlei He

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Tsinghua University(清华大学)

AI总结 RePPL通过校准语义传播和语言生成中的不确定性,提升问答幻觉检测性能,生成token级不确定性评分作为解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01156 2026-02-03 cs.LG cs.RO

PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning

PolicyFlow: 在强化学习中使用连续归一化流进行策略优化

Shunpeng Yang, Ben Liu, Hua Chen

机构 * Hong Kong University of Science and Technology(香港科技大学) Southern University of Science and Technology(南方科技大学) Zhejiang University-University of Illinois Urbana-Champaign Institute(浙江大学-伊利诺伊大学厄巴纳-香槟分校联合研究所) LimX Dynamics

AI总结 PolicyFlow是一种基于连续归一化流的在线强化学习算法,通过减少似然计算开销实现稳定训练,并通过布朗正则化器促进多样化行为。

Comments Submitted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00969 2026-02-03 cs.LG

On the Spectral Flattening of Quantized Embeddings

量化嵌入谱扁平化研究

Junlin Huang, Wenyi Fang, Zhenheng Tang, Yuxin Wang, Xueze Kang, Yang Zheng, Bo Li, Xiaowen Chu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Huawei Technologies Co., Ltd(华为技术有限公司)

AI总结 本研究揭示了量化嵌入谱扁平化现象,证明了谱保真度对稳定低比特优化的必要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00808 2026-02-03 cs.RO

Physics-informed Diffusion Mamba Transformer for Real-world Driving

物理引导的扩散Mamba变换器用于现实驾驶

Hang Zhou, Qiang Zhang, Peiran Liu, Yihao Qin, Zhaoxu Yan, Yiding Ji

机构 * Robotics and Autonomous Systems Thrust, The Hong Kong University of Science and Technology(Guangzhou)(香港科学与技术大学(广州)机器人与自主系统方向) MoSense Technologies

AI总结 本文提出了一种结合Mamba和注意力机制的扩散模型,通过整合物理约束提升自动驾驶轨迹预测的准确性和可解释性。

Journal ref International Conference on Robotics & Automation (ICRA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00739 2026-02-03 cs.CV

Diffusion-Driven Inter-Outer Surface Separation for Point Clouds with Open Boundaries

基于扩散的点云双层表面分离:用于开放边界

Zhengyan Qin, Liyuan Qiu

机构 * Hong Kong University of Science and Technology (HKUST)(香港理工大学) Hong Kong Applied Science and Technology Research Institute (ASTRI)(香港应用科技研究院)

AI总结 本文提出了一种基于扩散的算法,用于分离双层点云的内层和外层表面,特别针对具有开放边界的点云,通过提取真实内层来解决重叠表面和法线紊乱问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00726 2026-02-03 cs.HC cs.AI

Augmenting Clinical Decision-Making with an Interactive and Interpretable AI Copilot: A Real-World User Study with Clinicians in Nephrology and Obstetrics

通过交互式和可解释的AI助手增强临床决策:与泌尿科和产科医生的现实世界用户研究

Yinghao Zhu, Dehao Sui, Zixiang Wang, Xuning Hu, Lei Gu, Yifan Qi, Tianchen Wu, Ling Wang, Yuan Wei, Wen Tang, Zhihan Cui, Yasha Wang, Lequan Yu, Ewen M Harrison, Junyi Gao, Liantao Ma

机构 * Peking University(北京大学) University of Hong Kong(香港大学) Hong Kong University of Science and Technology(香港科学与技术大学) Peking University Third Hospital(北京大学第三医院) Affiliated Xuzhou Municipal Hospital of Xuzhou Medical University(徐州医科大学附属徐州市人民医院) University of Edinburgh(爱丁堡大学) Health Data Research UK(英国健康数据研究)

AI总结 AICare通过交互式和可解释的AI助手提升临床决策,通过实验证明其降低认知负荷并增强医生信任

Comments Accepted by ACM CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00686 2026-02-03 cs.RO

Learning to Accelerate Vision-Language-Action Models through Adaptive Visual Token Caching

通过自适应视觉令牌缓存学习加速视觉-语言-动作模型

Yujie Wei, Jiahan Fan, Jiyu Guo, Ruichen Zhen, Rui Shao, Xiu Su, Zeke Xie, Shuo Yang

机构 * Harbin Institute of Technology(哈尔滨工业大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Meituan Academy of Robotics Shenzhen, Meituan(美团机器人深圳研究院) Central South University(中南大学) HKUST(GZ)(香港科技大学(广州))

AI总结 本文提出通过自适应视觉令牌缓存学习加速VLA模型,提升推理效率并提高任务成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00528 2026-02-03 cs.AI

How Far Are LLMs from Professional Poker Players? Revisiting Game-Theoretic Reasoning with Agentic Tool Use

LLMs距离专业扑克玩家还有多远?结合代理工具使用的博弈论推理再探

Minhua Lin, Enyan Dai, Hui Liu, Xianfeng Tang, Yuliang Yan, Zhenwei Dai, Jingying Zeng, Zhiwei Zhang, Fali Wang, Hongcheng Gao, Chen Luo, Xiang Zhang, Qi He, Suhang Wang

机构 * The Pennsylvania State University(宾夕法尼亚州立大学) HKUST (GZ)(香港科技大学) Amazon(亚马逊) Tsinghua University(清华大学) Microsoft(微软公司)

AI总结 本文提出ToolPoker框架,通过整合外部求解器和专业解释,提升LLMs在扑克博弈中的推理和游戏表现。

Comments Accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20327 2026-02-03 cs.CL

CE-RM: A Pointwise Generative Reward Model Optimized via Two-Stage Rollout and Unified Criteria

CE-RM:一种通过两阶段回放和统一标准优化的点wise生成奖励模型

Xinyu Hu, Yancheng He, Weixun Wang, Tao Feng, Li Lin, Jiashun Liu, Wenbo Su, Bo Zheng, Xiaojun Wan

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学计算机技术研究院) Hong Kong University of Science and Technology(香港理工大学) Alibaba Group(阿里巴巴集团)

AI总结 CE-RM通过两阶段回放和统一标准优化,提升生成奖励模型在开放性自然语言生成和强化学习中的表现。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15248 2026-02-03 cs.LG cs.AI

EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control

EntroPIC: 通过比例-积分控制实现LLM长期训练的稳定性

Kai Yang, Xin Xu, Yangkun Chen, Weijie Liu, Jiafei Lyu, Zichuan Lin, Deheng Ye, Saiyong Yang

机构 * Tencent Hunyuan(腾讯文言) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 EntroPIC通过比例-积分控制实现LLM长期训练的熵稳定,提升探索效率和训练稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏