arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The University of Hong Kong(香港大学)

共收录 1522
2506.06440 2025-06-10 cs.GR cs.CV

Vid2Sim: Generalizable, Video-based Reconstruction of Appearance, Geometry and Physics for Mesh-free Simulation

Chuhao Chen, Zhiyang Dou, Chen Wang, Yiming Huang, Anjun Chen, Qiao Feng, Jiatao Gu, Lingjie Liu

机构 * University of Pennsylvania(宾夕法尼亚大学) The University of Hong Kong(香港大学) Zhejiang University(浙江大学)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06412 2025-06-10 cs.LG cs.CV

NeurNCD: Novel Class Discovery via Implicit Neural Representation

Junming Wang, Yi Shi

机构 * The University of Hong Kong(香港大学) Beijing Jiaotong University(北京交通大学)

Comments Accepted by ICMR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08468 2025-06-10 cs.RO cs.CV

Multi-GraspLLM: A Multimodal LLM for Multi-Hand Semantic Guided Grasp Generation

Haosheng Li, Weixin Mao, Weipeng Deng, Chenyu Meng, Haoqiang Fan, Tiancai Wang, Yoshie Osamu, Ping Tan, Hongan Wang, Xiaoming Deng

机构 * Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) Waseda University(早稻田大学) University of Hong Kong(香港大学) MEGVII Technology(美格智能科技) Hong Kong University of Science and Technology(香港科技大学)

Comments 16 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05871 2025-06-09 cs.LG cs.DC cs.PF

BestServe: Serving Strategies with Optimal Goodput in Collocation and Disaggregation Architectures

Xiannan Hu, Tianyou Zeng, Xiaoming Yuan, Liwei Song, Guangyuan Zhang, Bangzheng He

机构 * The University of Hong Kong(香港大学) Huawei Cloud, Huawei Technologies Co., Ltd(华为云、华为技术有限公司)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24870 2025-06-09 cs.CV

GenSpace: Benchmarking Spatially-Aware Image Generation

Zehan Wang, Jiayang Xu, Ziang Zhang, Tianyu Pang, Chao Du, Hengshuang Zhao, Zhou Zhao

机构 * Zhejiang University(浙江大学) Sea AI Lab(海思人工智能实验室) The University of Hong Kong(香港大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06805 2025-06-09 cs.LG cs.GR

Efficient Diffusion Models: A Survey

Hui Shen, Jingxuan Zhang, Boning Xiong, Rui Hu, Shoufa Chen, Zhongwei Wan, Xin Wang, Yu Zhang, Zixuan Gong, Guangyin Bao, Chaofan Tao, Yongfeng Huang, Ye Yuan, Mi Zhang

机构 * The Ohio State University(俄亥俄州立大学) Indiana University(印第安纳大学) Fudan University(复旦大学) Hangzhou City University(杭州城市大学) The University of Hong Kong(香港大学) Tongji University(同济大学) The Chinese University of Hong Kong(香港中文大学) Peking University(北京大学)

Comments Published in Transactions on Machine Learning Research (TMLR-2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.13926 2025-06-09 cs.LG

Quantifying the Optimization and Generalization Advantages of Graph Neural Networks Over Multilayer Perceptrons

Wei Huang, Yuan Cao, Haonan Wang, Xin Cao, Taiji Suzuki

机构 * RIKEN AIP(日本学术振兴会先进研究所) The University of Hong Kong(香港大学) National University of Singapore(新加坡国立大学) The University of New South Wales(新南威尔士大学) University of Tokyo(东京大学)

Journal ref AISTATS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05302 2025-06-06 cs.CV

Perceive Anything: Recognize, Explain, Caption, and Segment Anything in Images and Videos

Weifeng Lin, Xinyu Wei, Ruichuan An, Tianhe Ren, Tingwei Chen, Renrui Zhang, Ziyu Guo, Wentao Zhang, Lei Zhang, Hongsheng Li

机构 * CUHK(香港中文大学) HKU(香港大学) PolyU Peking University(北京大学)

Comments 19 pages, 13 figures, Website: https://Perceive-Anything.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04421 2025-06-06 cs.CV cs.AI cs.LG

HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation

Hermann Kumbong, Xian Liu, Tsung-Yi Lin, Ming-Yu Liu, Xihui Liu, Ziwei Liu, Daniel Y. Fu, Christopher Ré, David W. Romero

机构 * Stanford University(斯坦福大学) NVIDIA CUHK(中国香港大学) HKU(香港大学) NTU(国立新加坡大学) UCSD(加州大学圣地亚哥分校) Together AI

Comments Accepted to CVPR 2025. Project Page: https://research.nvidia.com/labs/dir/hmar/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04108 2025-06-06 cs.CL

Rectified Sparse Attention

Yutao Sun, Tianzhu Ye, Li Dong, Yuqing Xia, Jian Chen, Yizhao Gao, Shijie Cao, Jianyong Wang, Furu Wei

机构 * Microsoft Research(微软研究院) Tsinghua University(清华大学) The University of Hong Kong(香港大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04228 2025-06-05 cs.CV

LayerFlow: A Unified Model for Layer-aware Video Generation

Sihui Ji, Hao Luo, Xi Chen, Yuanpeng Tu, Yiyang Wang, Hengshuang Zhao

机构 * The University of Hong Kong(香港大学) DAMO Academy, Alibaba Group(阿里云达摩院) Alibaba Group(阿里巴巴集团) Hupan Laboratory(虎跑实验室)

Comments Project Page: https://sihuiji.github.io/LayerFlow-Page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04209 2025-06-05 cs.CV

Language-Image Alignment with Fixed Text Encoders

Jingfeng Yang, Ziyang Wu, Yue Zhao, Yi Ma

机构 * UC Berkeley(伯克利大学) The University of Hong Kong(香港大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03790 2025-06-05 cs.LG

Attention-Only Transformers via Unrolled Subspace Denoising

Peng Wang, Yifu Lu, Yaodong Yu, Druv Pai, Qing Qu, Yi Ma

机构 * Department of Electrical Engineering and Computer Science, University of Michigan, Ann Arbor(电气工程与计算机科学系,密歇根大学,安阿伯分校) Department of Electrical Engineering and Computer Science, University of California, Berkeley(电气工程与计算机科学系,加州大学伯克利分校) Institute of Data Science & School of Computing and Data Science, University of Hong Kong(数据科学研究所及计算与数据科学学院,香港大学)

Comments 28 pages, 7 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.13101 2025-06-05 cs.LG cs.AI cs.DC

AdaptSFL: Adaptive Split Federated Learning in Resource-constrained Edge Networks

Zheng Lin, Guanqiao Qu, Wei Wei, Xianhao Chen, Kin K. Leung

机构 * Department of Electrical and Electronic Engineering, University of Hong Kong(香港大学电子与电气工程系)

Comments 16 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16739 2025-06-05 cs.LG cs.AI

Pushing Large Language Models to the 6G Edge: Vision, Challenges, and Opportunities

Zheng Lin, Guanqiao Qu, Qiyuan Chen, Xianhao Chen, Zhe Chen, Kaibin Huang

机构 * Department of Electrical and Electronic Engineering, University of Hong Kong(香港大学电子与电气工程系) HKU Musketeers Foundation Institute of Data Science, University of Hong Kong(香港大学穆斯奎特基金会数据科学研究所) School of Computer Science, Fudan University(复旦大学计算机学院)

Comments 7 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03126 2025-06-04 cs.CV

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Lu Qiu, Yizhuo Li, Yuying Ge, Yixiao Ge, Ying Shan, Xihui Liu

机构 * The University of Hong Kong(香港大学) ARC Lab, Tencent PCG(腾讯PCG ARC实验室)

Comments Project released at: https://qiulu66.github.io/animeshooter/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02975 2025-06-04 cs.CV cs.AI

HaploOmni: Unified Single Transformer for Multimodal Video Understanding and Generation

Yicheng Xiao, Lin Song, Rui Yang, Cheng Cheng, Zunnan Xu, Zhaoyang Zhang, Yixiao Ge, Xiu Li, Ying Shan

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) ARC Lab, Tencent PCG(腾讯PCG ARC实验室) The University of Hong Kong(香港大学) Xi’an JiaoTong University(西安交通大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20129 2025-06-04 cs.CL cs.LG

Finite State Automata Inside Transformers with Chain-of-Thought: A Mechanistic Study on State Tracking

Yifan Zhang, Wenyu Du, Dongming Jin, Jie Fu, Zhi Jin

机构 * Key Laboratory of High Confidence Software Technology (PKU), MOE, China(高性能软件技术关键实验室(PKU),教育部,中国) School of Computer Science, Peking University, China(北京大学计算机学院,中国) The University of Hong Kong(香港大学) Shanghai AI Lab(上海人工智能实验室)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.15806 2025-06-04 cs.CR cs.AI cs.CL cs.LG

A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos

Yang Yao, Xuan Tong, Ruofan Wang, Yixu Wang, Lujundong Li, Liang Liu, Yan Teng, Yingchun Wang

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) The University of Hong Kong(香港大学) Fudan University(复旦大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00817 2025-06-03 cs.CL

One for All: Update Parameterized Knowledge Across Multiple Models

Weitao Ma, Xiyuan Du, Xiaocheng Feng, Lei Huang, Yichong Huang, Huiyi Zhang, Xiaoliang Yang, Baohang Li, Xiachong Feng, Ting Liu, Bing Qin

机构 * Harbin Institute of Technology(哈尔滨工业大学) Peng Cheng Laboratory(鹏城实验室) The University of Hong Kong(香港大学)

Comments ACL 2025 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00564 2025-06-03 eess.IV cs.CV

Image Restoration Learning via Noisy Supervision in the Fourier Domain

Haosen Liu, Jiahao Liu, Shan Tan, Edmund Y. Lam

机构 * Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系) Key Laboratory of Image Processing and Intelligent Control, Ministry of Education(教育部图像处理与智能控制重点实验室;华中科技大学人工智能与自动化学院) School of Artificial Intelligence and Automation, Huazhong University of Science and Technology

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18557 2025-06-03 cs.CL

TAG-INSTRUCT: Controlled Instruction Complexity Enhancement through Structure-based Augmentation

He Zhu, Zhiwen Ruan, Junyou Su, Xingwei He, Yun Chen, Wenjia Zhang, Guanhua Chen

机构 * Peking University(北京大学) Southern University of Science and Technology(南方科技大学) Tongji University(同济大学) Shanghai University of Finance and Economics(上海财经大学) The University of Hong Kong(香港大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14814 2025-06-03 cs.RO

VB-Com: Learning Vision-Blind Composite Humanoid Locomotion Against Deficient Perception

Junli Ren, Tao Huang, Huayi Wang, Zirui Wang, Qingwei Ben, Junfeng Long, Yanchao Yang, Jiangmiao Pang, Ping Luo

机构 * Shanghai AI Laboratory(上海人工智能实验室) The University of Hong Kong(香港大学) Shanghai Jiao Tong University(上海交通大学) Zhejiang University(浙江大学) The Chinese University of Hong Kong(香港中文大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17451 2025-06-03 cs.CV cs.CL

VL-RewardBench: A Challenging Benchmark for Vision-Language Generative Reward Models

Lei Li, Yuancheng Wei, Zhihui Xie, Xuqing Yang, Yifan Song, Peiyi Wang, Chenxin An, Tianyu Liu, Sujian Li, Bill Yuchen Lin, Lingpeng Kong, Qi Liu

机构 * HKU(香港大学) SCUT(华南理工大学) SJTU(上海交通大学) PKU(北京大学) Allen AI(AllenAI)

Comments CVPR 2025 Camera Ready Version. Project page: https://vl-rewardbench.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05902 2025-06-03 cs.CV cs.CL

Autoregressive Models in Vision: A Survey

Jing Xiong, Gongye Liu, Lun Huang, Chengyue Wu, Taiqiang Wu, Yao Mu, Yuan Yao, Hui Shen, Zhongwei Wan, Jinfa Huang, Chaofan Tao, Shen Yan, Huaxiu Yao, Lingpeng Kong, Hongxia Yang, Mi Zhang, Guillermo Sapiro, Jiebo Luo, Ping Luo, Ngai Wong

机构 * The University of Hong Kong(香港大学) Tsinghua University(清华大学) Duke University(杜克大学) University of Rochester(罗切斯特大学) The Ohio State University(俄亥俄州立大学) Bytedance(字节跳动) The University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) Apple(苹果公司) The Hong Kong Polytechnic University(香港理工大学) Princeton University(普林斯顿大学)

Comments The paper is accepted by TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08665 2025-06-03 cs.RO cs.SY eess.SY

Agile Decision-Making and Safety-Critical Motion Planning for Emergency Autonomous Vehicles

Yiming Shu, Jingyuan Zhou, Fu Zhang

机构 * University of Hong Kong, Department of Mechanical Engineering(香港大学机械工程系) National University of Singapore, Department of Civil and Environmental Engineering(新加坡国立大学土木与环境工程系)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00391 2025-06-03 cs.CL

SHARE: An SLM-based Hierarchical Action CorREction Assistant for Text-to-SQL

Ge Qu, Jinyang Li, Bowen Qin, Xiaolong Li, Nan Huo, Chenhao Ma, Reynold Cheng

机构 * The University of Hong Kong(香港大学) BAAI(百度人工智能研究院) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

Comments Accepted to ACL 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00173 2025-06-03 cs.GR cs.RO

MotionPersona: Characteristics-aware Locomotion Control

Mingyi Shi, Wei Liu, Jidong Mei, Wangpok Tse, Rui Chen, Xuelin Chen, Taku Komura

机构 * The University of Hong Kong(香港大学) Shandong University(山东大学) Hong Kong University of Science and Technology(香港科技大学) Adobe Research

Comments 15 pages, 13 figures, webpage: https://motionpersona25.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01710 2025-06-03 cs.CV cs.LG cs.RO

Enhancing Large Vision Model in Street Scene Semantic Understanding through Leveraging Posterior Optimization Trajectory

Wei-Bin Kou, Qingfeng Lin, Ming Tang, Jingreng Lei, Shuai Wang, Rongguang Ye, Guangxu Zhu, Yik-Chung Wu

机构 * Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系) Shenzhen Research Institute of Big Data(大数据研究深圳研究所) Department of Computer Science and Engineering, Southern University of Science and Technology(南方科技大学计算机科学与工程系) Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究所)

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17891 2025-06-03 cs.CL

Scaling Diffusion Language Models via Adaptation from Autoregressive Models

Shansan Gong, Shivam Agarwal, Yizhe Zhang, Jiacheng Ye, Lin Zheng, Mukai Li, Chenxin An, Peilin Zhao, Wei Bi, Jiawei Han, Hao Peng, Lingpeng Kong

机构 * The University of Hong Kong(香港大学) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Apple(苹果公司) Tencent AI Lab(腾讯AI实验室)

Comments ICLR 2025. (minor updates) Code: https://github.com/HKUNLP/DiffuLLaMA

详情

展开后加载摘要…

URL PDF HTML 收藏