arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

2026-03-27 至 2026-03-27 共收录 15
2603.25580 2026-03-27 cs.CV

UNIC: Neural Garment Deformation Field for Real-time Clothed Character Animation

UNIC:基于神经变形场的神经服装变形用于实时穿衣角色动画

Chengfeng Zhao, Junbo Qi, Yulou Liu, Zhiyang Dou, Minchen Li, Taku Komura, Ziwei Liu, Wenping Wang, Yuan Liu

机构 * HKUST(香港科技大学) Waseda(早稻田大学) MIT(麻省理工学院) CMU(卡内基梅隆大学) HKU(香港大学) NTU(南洋理工大学) TAMU(德克萨斯农工大学)

AI总结 本文提出UNIC方法,通过实例特定的神经变形场实时动画角色服装,无需泛化到新服装,提升变形质量和训练效率。

Comments Project page: https://igl-hkust.github.io/UNIC/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25565 2026-03-27 cs.CV

GeoHeight-Bench: Towards Height-Aware Multimodal Reasoning in Remote Sensing

GeoHeight-Bench:迈向遥感中基于高度的多模态推理

Xuran Hu, Zhitong Xiong, Zhongcheng Hong, Yifang Ban, Xiaoxiang Zhu, Wufan Zhao

机构 * KTH Royal Institute of Technology(瑞典皇家理工学院) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Technical University of Munich(慕尼黑工业大学)

AI总结 本文提出GeoHeight-Bench,通过构建高度感知的遥感理解评估框架,解决现有多模态模型在复杂遥感几何和灾害场景中忽略垂直维度的问题,提出数据生成管道和基准测试,并验证高度感知的重要性。

Comments 18 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25412 2026-03-27 cs.AI cs.CR

Beyond Content Safety: Real-Time Monitoring for Reasoning Vulnerabilities in Large Language Models

超越内容安全:大型语言模型推理漏洞的实时监控

Xunguang Wang, Yuguang Zhou, Qingyue Wang, Zongjie Li, Ruixuan Huang, Zhenlan Ji, Pingchuan Ma, Shuai Wang

机构 * The Hong Kong University of Science and Technology(香港科技大学) Zhejiang University of Technology(浙江工业大学)

AI总结 本文提出实时监控方法,识别大型语言模型推理过程中的安全漏洞,定义九类不安全推理行为,并通过实验验证其有效性,展示了推理安全的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25357 2026-03-27 cs.CV

InstanceAnimator: Multi-Instance Sketch Video Colorization

Yinhan Zhang, Yue Ma, Bingyuan Wang, Kunyu Feng, Yeying Jin, Qifeng Chen, Anyi Rao, Zeyu Wang

机构 * HKUST(GZ)(香港科技大学(广州)) HKUST(香港科技大学) NUS(新加坡国立大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25196 2026-03-27 cs.CL cs.AI

A Decade-Scale Benchmark Evaluating LLMs' Clinical Practice Guidelines Detection and Adherence in Multi-turn Conversations

一个十年级基准评估LLM在多轮对话中检测和遵守临床实践指南的能力

Andong Tan, Shuyu Dai, Jinglu Wang, Fengtao Zhou, Yan Lu, Xi Wang, Yingcong Chen, Can Yang, Shujie Liu, Hao Chen

机构 * Microsoft Research Asia(微软亚洲研究院) Hong Kong University of Science and Technology(香港科技大学) Peking University(北京大学) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 本文提出CPGBench基准,评估LLM在多轮对话中检测和遵循临床实践指南的能力,发现多数模型在识别和遵循指南方面存在显著差距,通过人类评估验证了这一发现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.25131 2026-03-27 cs.CV

Denoise and Align: Towards Source-Free UDA for Robust Panoramic Semantic Segmentation

去噪与对齐:迈向无源无监督领域适应的鲁棒全景语义分割

Yaowen Chang, Zhen Cao, Xu Zheng, Xiaoxin Mi, Zhen Dong

机构 * Wuhan University(武汉大学) HKUST (Guangzhou)(香港科技大学(广州)) Wuhan University of Technology(武汉理工大学)

AI总结 本文提出DAPASS框架,通过全景置信度引导去噪和上下文分辨率对抗模块,解决无源领域适应中因几何失真和数据缺失导致的伪标签不准确问题,提升户外和室内场景的语义分割性能。

Comments Accepted to CVPR26

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23997 2026-03-27 cs.CV

HGGT: Robust and Flexible 3D Hand Mesh Reconstruction from Uncalibrated Images

HGGT:从未校准图像中鲁棒且灵活的3D手形重建

Yumeng Liu, Xiao-Xiao Long, Marc Habermann, Xuanze Yang, Cheng Lin, Yuan Liu, Yuexin Ma, Wenping Wang, Ligang Liu

机构 * University of Science and Technology of China(中国科学技术大学) Nanjing University(南京大学) Max-Planck-Institut für Informatik(马克斯·普朗克信息学研究所) Macau University of Science and Technology(澳门科技大学) Hong Kong University of Science and Technology(香港科技大学) ShanghaiTech University(上海科技大学)

AI总结 本文提出HGGT方法,通过将手形重建视为视觉-几何基础任务,首次联合推断3D手形和相机姿态,实现了从未校准视角的鲁棒性和灵活性,优于现有方法并在真实场景中表现优异。

Comments project page: https://lym29.github.io/HGGT/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.22883 2026-03-27 cs.CV

Group Editing: Edit Multiple Images in One Go

组级编辑:一次操作编辑多张图像

Yue Ma, Xinyu Wang, Qianli Ma, Qinghe Wang, Mingzhe Zheng, Xiangpeng Yang, Hao Li, Chongbo Zhao, Jixuan Ying, Harry Yang, Hongyu Liu, Qifeng Chen

机构 * HKUST(香港科技大学) THU(清华大学) SJTU(上海交通大学) University of Technology Sydney(悉尼科技大学)

AI总结 本文提出组级编辑框架,通过显式和隐式关系建立实现多图像一致修改,引入融合机制和新数据集提升编辑质量与一致性。

Comments Accepted by CVPR 2026, Project page: https://group-editing.github.io/, Github: https://github.com/mayuelala/GroupEditing

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05042 2026-03-27 cs.CV cs.RO

CoIn3D: Revisiting Configuration-Invariant Multi-Camera 3D Object Detection

CoIn3D:重新审视配置不变的多摄像头3D目标检测

Zhaonian Kuang, Rui Ding, Haotian Wang, Xinhu Zheng, Meng Yang, Gang Hua

机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能全国重点实验室) Intelligent Transportation Thrust of the Systems Hub, HKUST(GZ)(香港科技大学(广州)系统枢纽智能交通学域) Amazon Alexa AI(亚马逊Alexa人工智能)

AI总结 本文提出CoIn3D框架,通过空间先验整合和摄像头感知数据增强,提升多摄像头3D目标检测在不同配置下的泛化能力。

Comments Accepted to CVPR 2026 main track

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24974 2026-03-27 math.OC cs.LG stat.ML

The Value of Information in Resource-Constrained Pricing

信息价值在资源受限定价中的作用

Ruicheng Ao, Jiashuo Jiang, David Simchi-Levi

机构 * Institute for Data, Systems, and Society, Massachusetts Institute of Technology(麻省理工学院数据、系统与社会研究所) Department of Civil and Environmental Engineering and Operations Research Center, MIT(麻省理工学院土木与环境工程系及运筹学研究中心) Department of Industrial Engineering and Decision Analytics, Hong Kong University of Science and Technology(香港科技大学工业工程与决策分析系)

AI总结 本文研究了在资源受限条件下,预测不确定性如何影响动态定价决策,通过线性需求、随机噪声和有限容量,证明了预测误差阈值对 regret 的影响,并展示了代理模型在降低方差中的作用。

Comments Extended version of the NeurIPS 2025 paper (arXiv:2501.14155). This version adds phase transition, surrogate-assisted variance reduction under model misspecification, and numerical experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24876 2026-03-27 cs.CV

OptiSAR-Net++: A Large-Scale Benchmark and Transformer-Free Framework for Cross-Domain Remote Sensing Visual Grounding

OptiSAR-Net++:跨领域遥感视觉定位的大型基准和无Transformer框架

Xiaoyu Tang, Jun Dong, Jintao Cheng, Rui Fan

机构 * South China Normal University(华南师范大学) Hong Kong University of Science and Technology(香港科技大学) Tongji University(同济大学) Shanghai Research Institute for Intelligent Autonomous Systems(上海自主智能无人系统科学中心) State Key Laboratory of Intelligent Autonomous Systems(智能自主系统国家重点实验室) Frontiers Science Center for Intelligent Autonomous Systems(智能自主系统前沿科学中心)

AI总结 本文提出OptiSAR-Net++,通过构建首个跨领域遥感视觉定位基准数据集,解决跨领域特征建模、计算效率和细粒度语义区分问题,提升定位精度与效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.24278 2026-03-27 cs.CV

TopoMesh: High-Fidelity Mesh Autoencoding via Topological Unification

TopoMesh: 通过拓扑统一实现高保真的网格自编码

Guan Luo, Xiu Li, Rui Chen, Xuanyu Yi, Jing Lin, Chia-Hao Chen, Jiahang Liu, Song-Hai Zhang, Jianfeng Zhang

机构 * Tsinghua University(清华大学) ByteDance Seed(字节跳动种子) HKUST(香港科技大学)

AI总结 TopoMesh通过拓扑统一框架解决网格生成中的拓扑不匹配问题,利用稀疏体素VAE实现高保真重建,提升几何细节和锐利特征的保留能力。

Comments Accepted to CVPR 2026. Project page: https://logan0601.github.io/projects/topomesh/index.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01939 2026-03-27 cs.RO cs.AI

Towards Exploratory and Focused Manipulation with Bimanual Active Perception: A New Problem, Benchmark and Strategy

迈向探索性与聚焦性操控的双臂主动感知:一个新的问题、基准和策略

Yuxin He, Ruihao Zhang, Tianao Shen, Cheng Liu, Qiang Nie

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 本文提出探索性与聚焦性操控(EFM)问题,建立EFM-10基准并提出双臂主动感知策略,通过主动视觉和力觉感知提升复杂操控任务的完成能力。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16889 2026-03-27 cs.CL

Can GRPO Boost Complex Multimodal Table Understanding?

GRPO能否提升复杂多模态表格理解?

Xiaoqiang Kang, Shengen Wu, Zimu Wang, Yilin Liu, Xiaobo Jin, Kaizhu Huang, Wei Wang, Yutao Yue, Xiaowei Huang, Qiufeng Wang

机构 * School of Advanced Technology, Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学先进科技学院) Department of Computer Science, University of Liverpool(利物浦大学计算机科学系) Information Hub, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)信息中心) University of Southern California(南加州大学) Duke Kunshan University(杜克大学昆山分校)

AI总结 本文提出Table-R1框架,通过预热、感知对齐GRPO和提示-完成GRPO三阶段提升多模态表格理解性能,优于SFT和GRPO。

Comments EMNLP 2025

Journal ref EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.10028 2026-03-27 cs.CV cs.RO

3D Dynamics-Aware Manipulation: Endowing Manipulation Policies with 3D Foresight

3D动态感知操控:赋予操控策略三维前瞻性

Yuxin He, Ruihao Zhang, Xianzu Wu, Zhiyuan Zhang, Cheng Ding, Qiang Nie

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) JAKA Robotics Co., Ltd.(JAKA机器人有限公司)

AI总结 本文提出一种融合3D世界建模与策略学习的框架,通过引入三个自监督任务提升操控策略的3D前瞻性,实验表明在仿真和现实环境中显著提升操控性能。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏