arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

共收录 2801
2512.10957 2025-12-12 cs.CV cs.AI

SceneMaker: Open-set 3D Scene Generation with Decoupled De-occlusion and Pose Estimation Model

SceneMaker: 一种解耦的开放集3D场景生成方法,结合去遮挡与姿态估计模型

Yukai Shi, Weiyu Li, Zihao Wang, Hongyang Li, Xingyu Chen, Ping Tan, Lei Zhang

机构 * Tsinghua University(清华大学) HKUST(香港科技大学) IDEA Research(IDEA研究院) LightIllusions

AI总结 SceneMaker通过解耦去遮挡与姿态估计模型,提升开放集3D场景生成的准确性和泛化能力。

Comments Project page: https://idea-research.github.io/SceneMaker/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10428 2025-12-12 cs.ET cs.RO

Design and Implementation of a High-Precision Wind-Estimation UAV with Onboard Sensors

高精度风速估计无人机的 design 和实现

Haowen Yu, Na Fan, Xing Liu, Ximin Lyu

机构 * School of Intelligent Systems Engineering, Sun Yat-sen University(中山大学智能系统工程学院) Department of Computer Science and Engineering, Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系) Shenzhen ZEEY Technology Co., Ltd.(深圳ZEEY科技有限公司)

AI总结 本文提出了一种基于机载传感器的高精度风速估计方法,通过扰动观测器和薄板样条模型实现高精度风向量估计,实验结果在多种场景下均优于现有基线。

Comments https://www.sciencedirect.com/science/article/abs/pii/S0263224125032415?via%3Dihub

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07752 2025-12-12 cs.CV

DEGS: Deformable Event-based 3D Gaussian Splatting from RGB and Event Stream

DEGS: 从RGB和事件流中基于可变形的3D高斯点云重建

Junhao He, Jiaxu Wang, Jia Li, Mingyuan Sun, Qiang Zhang, Jiahang Cao, Ziyi Zhang, Yi Gu, Jingkai Sun, Renjing Xu

机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Hong Kong University of Science and Technology(香港科技大学) University of Hong Kong(香港大学) Northeastern University(东北大学)

AI总结 本文提出DEGS框架,结合RGB和事件流优化动态3DGS,通过事件运动先验指导变形场优化,提升重建性能。

Comments Accepted by IEEE TVCG

Journal ref 2025 IEEE Transactions on Visualization and Computer Graphics

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04522 2025-12-12 cs.LG cs.AI

Toward a Unified Geometry Understanding: Riemannian Diffusion Framework for Graph Generation and Prediction

迈向统一几何理解:图生成与预测的黎曼扩散框架

Yisen Gao, Xingcheng Fu, Qingyun Sun, Jianxin Li, Xianxian Li

机构 * Key Lab of Education Blockchain and Intelligent Technology, Guangxi Normal University(教育区块链与智能技术重点实验室,广西师范大学) Computer Science and Engineering, The Hong Kong University of Science and Technology(计算机科学与工程,香港科学与技术大学) Guangxi Key Lab of Multi-source Information Mining & Security, Guangxi Normal University(广西多源信息挖掘与安全重点实验室,广西师范大学) School of Computer Science and Engineering, Beihang University(计算机科学与工程学院,北京航空航天大学)

AI总结 本文提出GeoMancer框架,通过黎曼扩散方法解决图数据生成与预测中的几何潜力释放问题,提升模型对复杂流形结构的学习能力。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01426 2025-12-12 cs.LG cs.AI

UniExtreme: A Universal Foundation Model for Extreme Weather Forecasting

UniExtreme: 一种用于极端天气预报的通用基础模型

Hang Ni, Weijia Zhang, Hao Liu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

AI总结 UniExtreme是一种通用极端天气预报基础模型,通过自适应频率调制和事件先验增强模块,提升对多样化极端天气事件的预测能力。

Comments 35 pages, 80 figures, submitted to ACM KDD 2026 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09936 2025-12-12 eess.SY cs.LG cs.SY quant-ph

QSTAformer: A Quantum-Enhanced Transformer for Robust Short-Term Voltage Stability Assessment against Adversarial Attacks

QSTAformer: 一种增强型变压器用于对抗攻击下的短期电压稳定性评估

Yang Li, Chong Ma, Yuanzheng Li, Sen Li, Yanbo Chen, Zhaoyang Dong

机构 * School of Electrical Engineering, Northeast Electric Power University(东北电力大学电气工程学院) State Grid Shandong Electric Power Company Jiaozhou Power Supply Company(山东电网济南供电分公司) School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院) Department of Civil and Environmental Engineering, The Hong Kong University of Science and Technology(香港科技大学土木与环境工程系) State Key Laboratory of Alternate Electrical Power System with Renewable Energy Sources, School of Electrical & Electronic Engineering, North China Electric Power University(可再生能源电力系统国家重点实验室,华北电力大学电气与电子工程学院) Department of Electrical Engineering, City University of Hong Kong(香港城市大学电气工程系)

AI总结 QSTAformer通过整合参数化量子电路的增强型变压器架构,提升对抗攻击下的短期电压稳定性评估的鲁棒性和效率。

Comments 15 pages, 12 figures. Accepted by Applied Energy

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09483 2025-12-11 cs.CL cs.CY

Source Coverage and Citation Bias in LLM-based vs. Traditional Search Engines

基于大语言模型的搜索引擎与传统搜索引擎的来源覆盖与引用偏见

Peixian Zhang, Qiming Ye, Zifan Peng, Kiran Garimella, Gareth Tyson

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Rutgers University(罗格斯大学) Rutgers University New Brunswick United States(罗格斯大学新 Brunswick美国)

AI总结 本文研究了基于大语言模型的搜索引擎与传统搜索引擎在来源覆盖和引用偏见方面的差异,发现LLM-SEs在资源多样性上优于传统搜索引擎,但其可信度和中立性仍需进一步提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09441 2025-12-11 cs.CV cs.AI

Representation Calibration and Uncertainty Guidance for Class-Incremental Learning based on Vision Language Model

基于视觉语言模型的类增量学习中的表示校准与不确定性引导

Jiantao Tan, Peixian Ma, Tong Yu, Wentao Zhang, Ruixuan Wang

机构 * Guangdong Province Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education(广东省机器智能与先进计算重点实验室) Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Peng Cheng Laboratory(鹏城实验室) Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education(广东省机器智能与先进计算重点实验室)

AI总结 本文提出了一种基于视觉语言模型的类增量学习框架,通过引入任务特定适配器和跨任务表示校准策略,提升类别区分能力,并利用预测不确定性优化图像特征选择,实现更准确的分类性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09267 2025-12-11 cs.LG

Contrastive Learning for Semi-Supervised Deep Regression with Generalized Ordinal Rankings from Spectral Seriation

基于谱序列化通用序排名的半监督深度回归对比学习

Ce Wang, Weihang Dai, Hanru Bai, Xiaomeng Li

机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学电子与计算机工程系) School of Science, Sun Yat-sen University(中山大学科学学院) Institute of Science and Technology for Brain-Inspired Intelligence, Fudan University(复旦大学脑启发式智能科学与技术研究院)

AI总结 本文提出一种基于谱序列化通用序排名的半监督深度回归对比学习方法,通过引入标记样本和未标记样本构建特征相似性矩阵,提升模型鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08294 2025-12-11 cs.CV

OpenSubject: Leveraging Video-Derived Identity and Diversity Priors for Subject-driven Image Generation and Manipulation

OpenSubject: 利用视频衍生的身份和多样性先验进行主体驱动的图像生成与操纵

Yexin Liu, Manyuan Zhang, Yueze Wang, Hongyu Li, Dian Zheng, Weiming Zhang, Changsheng Lu, Xunliang Cai, Yan Feng, Peng Pei, Harry Yang

机构 * HKUST(香港科技大学) Meituan(美团) HKUST(GZ)(香港科技大学(广州))

AI总结 OpenSubject通过视频衍生的大规模数据集提升主体驱动图像生成与操纵的性能,尤其在复杂场景中表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07596 2025-12-11 cs.CV cs.RO

More than Segmentation: Benchmarking SAM 3 for Segmentation, 3D Perception, and Reconstruction in Robotic Surgery

超越分割:在机器人手术中评估SAM 3用于分割、3D感知和重建的基准测试

Wenzhen Dong, Jieming Yu, Yiming Huang, Hongqiu Wang, Lei Zhu, Albert C. S. Chung, Hongliang Ren, Long Bai

机构 * The Chinese University of Hong Kong(香港中文大学) The Hong Kong University of Science and Technology(香港科学与技术大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Technical University of Munich(慕尼黑技术大学)

AI总结 本文评估了SAM 3在机器人手术中的分割、3D感知和重建能力,展示了其在动态视频跟踪中的有效性,同时指出了基于语言提示的局限性及进一步改进的必要性。

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08814 2025-12-10 cs.CL

Ask, Answer, and Detect: Role-Playing LLMs for Personality Detection with Question-Conditioned Mixture-of-Experts

提问、回答与检测:基于角色扮演的LLM用于带有问题条件的专家混合模型的人格检测

Yifan Lyu, Liang Zhang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) University of International Business and Economics(国际经济贸易大学)

AI总结 ROME通过角色扮演LLM模拟用户回答心理测量问卷,生成可解释的证据链接语言线索与人格标签,提升人格检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08358 2025-12-10 cs.CV

TrackingWorld: World-centric Monocular 3D Tracking of Almost All Pixels

TrackingWorld: 以世界为中心的单目3D像素跟踪

Jiahao Lu, Weitao Xiong, Jiacheng Deng, Peng Li, Tianyu Huang, Zhiyang Dou, Cheng Lin, Sai-Kit Yeung, Yuan Liu

机构 * HKUST(香港科技大学) USTC(中国科学技术大学) CUHK(香港中文大学) HKU(香港大学) XMU(厦门大学) MUST(澳门科技大学)

AI总结 TrackingWorld提出了一种以世界为中心的单目3D像素跟踪方法,通过上采样器和优化框架实现对视频中几乎所有像素的密集3D跟踪。

Comments Accepted by NeurIPS 2025. Project Page: https://igl-hkust.github.io/TrackingWorld.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08300 2025-12-10 cs.AI

rSIM: Incentivizing Reasoning Capabilities of LLMs via Reinforced Strategy Injection

rSIM: 通过强化策略注入激励大语言模型的推理能力

Sijia Chen, Baochun Li, Di Niu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) University of Toronto(多伦多大学) University of Alberta(阿尔伯塔大学)

AI总结 rSIM通过强化策略注入机制,使LLM具备推理能力,并在实验中显著提升模型性能。

Comments 14 pages, 6 figures. Accepted to the ACL ARR July

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06368 2025-12-10 cs.CV

HuPrior3R: Incorporating Human Priors for Better 3D Dynamic Reconstruction from Monocular Videos

HuPrior3R: 融合人类先验以实现从单目视频中更好的3D动态重建

Weitao Xiong, Zhiyuan Yuan, Jiahao Lu, Chengfeng Zhao, Peng Li, Yuan Liu

机构 * HKUST(香港科技大学) XMU(厦门大学) SYSU(南方科技大学)

AI总结 HuPrior3R通过融合SMPL人体模型与单目深度估计,提出混合几何先验方法,提升单目视频中动态人类3D重建的几何一致性和细节精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17114 2025-12-10 cs.AI

Mathematical Proof as a Litmus Test: Revealing Failure Modes of Advanced Large Reasoning Models

数学证明作为检验标准:揭示先进大推理模型的失败模式

Dadi Guo, Jiayu Liu, Zhiyuan Fan, Zhitao He, Haoran Li, Yuxin Li, Yumeng Wang, Yi R. Fung

机构 * Hong Kong University of Science and Technology(香港科技大学)

AI总结 本文通过引入RFMDataset揭示大推理模型在数学证明中的失败模式,发现其在严谨性和正确性上的不足,需进一步提升逻辑训练。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05872 2025-12-10 cs.CV

Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection

领域-RAG:跨领域少样本目标检测的检索引导组合图像生成

Yu Li, Xingyu Qiu, Yuqian Fu, Jie Chen, Tianwen Qian, Xu Zheng, Danda Pani Paudel, Yanwei Fu, Xuanjing Huang, Luc Van Gool, Yu-Gang Jiang

机构 * Fudan University(复旦大学) Fuzhou University(福州大学) East China Normal University(华东师范大学) HKUST(GZ)(香港科技大学(广州))

AI总结 Domain-RAG通过检索引导的组合图像生成方法,解决跨领域少样本目标检测中的领域对齐与背景生成问题,实现高质量样本生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18404 2025-12-10 cond-mat.stat-mech cs.LG

Deep generative modelling of canonical ensemble with differentiable thermal properties

基于可微热性质的 canonical 集合深度生成建模

Shuo-Hui Li, Yao-Wen Zhang, Ding Pan

机构 * Department of Physics, The Hong Kong University of Science and Technology(香港科技大学物理系) Department of Chemistry, The Hong Kong University of Science and Technology(香港科技大学化学系)

AI总结 本文提出了一种基于可微温度的变分方法,用于计算 canonical 集合的热力学量,实现了高效准确的直接采样模拟。

Comments Main text: 11.5 pages, 4 figures. Supplement: 20 pages. Github link: https://github.com/li012589/vatd

Journal ref Phys. Rev. Lett. 135, 027301 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07831 2025-12-09 cs.CV

UnityVideo: Unified Multi-Modal Multi-Task Learning for Enhancing World-Aware Video Generation

UnityVideo: 一体化多模态多任务学习以增强世界感知视频生成

Jiehui Huang, Yuechen Zhang, Xu He, Yuan Gao, Zhi Cen, Bin Xia, Yan Zhou, Xin Tao, Pengfei Wan, Jiaya Jia

机构 * HKUST(香港科技大学) CUHK(香港中文大学) Tsinghua University(清华大学) Kling Team, Kuaishou Technology(快手科技 Kling 团队)

AI总结 UnityVideo通过一体化多模态多任务学习提升视频生成的世界感知能力,引入动态噪声化和模态切换器,实现更全面的世界知识表示。

Comments Project Website https://jackailab.github.io/Projects/UnityVideo

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07464 2025-12-09 cs.RO

Gait-Adaptive Perceptive Humanoid Locomotion with Real-Time Under-Base Terrain Reconstruction

具有实时底层地形重建的感知人体机器人运动

Haolin Song, Hongbo Zhu, Tao Yu, Yan Liu, Mingqi Yuan, Wengang Zhou, Hua Chen, Houqiang Li

机构 * Department of Electronic Engineering and Information Science (EEIS), University of Science and Technology of China(电子工程与信息科学系,中国科学技术大学) LimX Dynamics(LimX动力学) Hong Kong University of Science and Technology(香港科技大学) School of Mechanics Engineering, Harbin Institute of Technology (HIT)(机械工程学院,哈尔滨工业大学) Department of Computing, The Hong Kong Polytechnic University(计算学院,香港理工大学) Zhejiang University-University of Illinois Urbana-Champaign Institute (ZJUI)(浙江大学-伊利诺伊大学厄巴纳-香槟分校联合研究所)

AI总结 该研究提出了一种结合地形感知、步态调节和全身控制的强化学习框架,通过实时底层地形重建实现稳健的人形机器人运动。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07385 2025-12-09 cs.CV

How Far are Modern Trackers from UAV-Anti-UAV? A Million-Scale Benchmark and New Baseline

现代追踪器距离无人机-反无人机还有多远?一个百万级基准和新基线

Chunhui Zhang, Li Liu, Zhipeng Zhang, Yong Wang, Hao Wen, Xi Zhou, Shiming Ge, Yanfeng Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) CloudWalk Technology Co., Ltd(云walk科技有限公司) School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) School of Aeronautics and Astronautics, Sun Yat-sen University(中山大学航空宇航学院) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)

AI总结 本文提出UAV-Anti-UAV多模态追踪任务及MambaSTS基线方法,通过百万级数据集验证了现有追踪算法在无人机反无人机领域的改进空间。

Comments https://github.com/983632847/Awesome-Multimodal-Object-Tracking

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07312 2025-12-09 cs.AR cs.AI cs.DC

DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management

DCO: 通过预测管理实现LLM加速器的动态缓存编排

Zhongchun Zhou, Chengtao Lai, Yuhang Gu, Wei Zhang

机构 * Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学) School of Electronic Science and Engineering, Southeast University(电子科学与工程学院,东南大学)

AI总结 DCO通过预测管理实现LLM加速器的动态缓存编排,利用数据流信息优化缓存替换和旁路决策,提升性能至1.8倍,面积仅0.064mm²。

Comments \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06999 2025-12-09 cs.SD cs.AI

Singing Timbre Popularity Assessment Based on Multimodal Large Foundation Model

基于多模态大基础模型的歌唱音色受欢迎程度评估

Zihao Wang, Ruibin Yuan, Ziqi Geng, Hengjia Li, Xingwei Qu, Xinyi Li, Songye Chen, Haoying Fu, Roger B. Dannenberg, Kejun Zhang

机构 * Zhejiang University(浙江大学) Carnegie Mellon University(卡内基梅隆大学) Hong Kong University of Science and Technology(香港科学与技术大学) University of California, Berkeley(加州大学伯克利分校) University of Manchester(曼彻斯特大学) Innovation Center of Yangtze River Delta, Zhejiang University(长江三角洲创新中心,浙江大学)

AI总结 本文提出基于多模态大基础模型的歌唱音色受欢迎程度评估方法,通过引入Sing-MD数据集、VocalVerse架构和H-TPR基准,实现无参考、多维度的歌唱评估。

Comments Accepted to ACMMM 2025 oral

Journal ref Proceedings of the 33rd ACM International Conference on Multimedia (ACMMM 2025), Pages 12227-12236

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19397 2025-12-09 cs.LG

Are Time-Series Foundation Models Deployment-Ready? A Systematic Study of Adversarial Robustness Across Domains

时间序列基础模型是否已准备好部署?跨领域的对抗鲁棒性系统研究

Jiawen Zhang, Zhenwei Zhang, Shun Zheng, Xumeng Wen, Jia Li, Jiang Bian

机构 * HKUST (GZ)(香港科技大学) Tsinghua University(清华大学) Microsoft Research Asia(微软亚洲研究院)

AI总结 本研究系统分析了时间序列基础模型在跨领域对抗攻击下的鲁棒性,揭示了模型脆弱性及改进方法。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.06272 2025-12-09 cs.CL cs.CE

Golden Touchstone: A Comprehensive Bilingual Benchmark for Evaluating Financial Large Language Models

黄金触点:一种全面的双语基准,用于评估金融大语言模型

Xiaojun Wu, Junxi Liu, Huanyi Su, Zhouchi Lin, Yiyan Qi, Chengjin Xu, Jiajun Su, Jiajie Zhong, Fuwei Wang, Saizhuo Wang, Fengrui Hua, Jia Li, Jian Guo

机构 * IDEA Research(IDEA研究机构) The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) The Hong Kong University of Science and Technology(香港科学与技术大学) Nanjing University(南京大学) South China Normal University(华南师范大学) DataArcTech Ltd.(DataArcTech有限公司)

AI总结 本研究提出Golden Touchstone双语基准,用于全面评估金融大语言模型的性能,通过对比分析揭示模型在处理复杂金融信息时的优势与局限。

Comments Published in Findings of EMNLP 2025

Journal ref In Findings of the Association for Computational Linguistics: EMNLP 2025, pages 22544-22560, Suzhou, China. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02483 2025-12-09 cs.CV

Event-Customized Image Generation

事件定制图像生成

Zhen Wang, Yilei Jiang, Dong Zheng, Jun Xiao, Long Chen

机构 * Zhejiang University, Hangzhou, China(浙江大学) The Hong Kong University of Science and Technology(香港科技大学)

AI总结 本文提出FreeEvent方法,通过引入实体切换和事件转移路径,实现事件定制化图像生成,提升复杂场景下的定制化能力。

Journal ref ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06866 2025-12-09 cs.CV cs.AI cs.CL cs.LG

Less Is More, but Where? Dynamic Token Compression via LLM-Guided Keyframe Prior

少即是多,但在哪里?通过LLM引导的关键帧先验实现动态令牌压缩

Yulin Li, Haokun Gui, Ziyang Fan, Junjie Wang, Bin Kang, Bin Chen, Zhuotao Tian

机构 * Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) Shenzhen Loop Area Institute(深圳河套学院) University of Chinese Academy of Sciences(中国科学院大学) The Hong Kong University of Science and Technology(香港科技大学)

AI总结 本文提出DyToK方法,通过LLM引导的关键帧先验实现动态令牌压缩,提升视频处理效率和准确性。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06854 2025-12-09 cs.AR cs.AI

ArchPower: Dataset for Architecture-Level Power Modeling of Modern CPU Design

ArchPower:现代CPU设计架构级功率建模的数据集

Qijun Zhang, Yao Lu, Mengming Li, Shang Liu, Zhiyao Xie

机构 * Hong Kong University of Science and Technology(香港理工大学)

AI总结 ArchPower是首个公开的架构级处理器功率建模数据集,通过复杂设计流程收集200个CPU样本,包含100+架构特征和细粒度功率信息,用于提升机器学习模型的准确性。

Comments Published in NeurIPS'25 Dataset and Benchmark Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06281 2025-12-09 cs.CV cs.AI

Unleashing the Intrinsic Visual Representation Capability of Multimodal Large Language Models

释放多模态大语言模型的内在视觉表示能力

Hengzhuang Li, Xinsong Zhang, Qiming Peng, Bin Luo, Han Hu, Dengyang Jiang, Han-Jia Ye, Teng Zhang, Hai Jin

机构 * HUST(华中科技大学) Tencent Hunyuan Research(腾讯混元研究) HKUST(香港科技大学) NJU(南京大学)

AI总结 本文提出LaVer框架,通过掩码图像建模提升多模态大语言模型的视觉表示能力,实验表明其在需要密集视觉能力的场景中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23880 2025-12-09 cs.CV cs.GR

TRELLISWorld: Training-Free World Generation from Object Generators

TRELLISWorld: 从物体生成器中无需训练的世界生成

Hanke Chen, Yuan Liu, Minchen Li

机构 * Carnegie Mellon University(卡内基梅隆大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 TRELLISWorld通过无需训练的模块化瓷砖生成器实现通用语言驱动的3D场景生成,支持多样布局、高效生成和灵活编辑。

详情

展开后加载摘要…

URL PDF HTML 收藏