arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Huazhong University of Science and Technology(华中科技大学)

共收录 702
2511.20410 2025-11-26 cs.CV

Image-Free Timestep Distillation via Continuous-Time Consistency with Trajectory-Sampled Pairs

无图像时间步蒸馏:通过连续时间一致性与轨迹采样对偶

Bao Tang, Shuai Zhang, Yueting Zhu, Jijun Xiang, Xin Yang, Li Yu, Wenyu Liu, Xinggang Wang

机构 * Huazhong University of Science and Technology(华中科技大学)

AI总结 TBCM通过直接提取教师模型生成轨迹的潜在表示,实现无训练数据依赖的时间步蒸馏,提升效率和简洁性,同时在生成质量与资源消耗上表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18416 2025-11-25 cs.CV

4D-VGGT: A General Foundation Model with SpatioTemporal Awareness for Dynamic Scene Geometry Estimation

4D-VGGT:一种具有时空意识的通用基础模型,用于动态场景几何估计

Haonan Wang, Hanyu Zhou, Haoyue Liu, Luxin Yan

机构 * National Key Lab of Multispectral Information Intelligent Processing Technology, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(多谱信息智能处理国家实验室,人工智能与自动化学院,华中科技大学) School of Computing, National University of Singapore(计算学院,新加坡国立大学)

AI总结 4D-VGGT通过分而治之的时空表示方法,提升动态场景几何估计的准确性和通用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18025 2025-11-25 cs.CR cs.IT cs.LG math.IT

Correlated-Sequence Differential Privacy

关联序列差分隐私

Yifan Luo, Meng Zhang, Jin Xu, Junting Chen, Jianwei Huang

机构 * Shenzhen Institute of Artificial Intelligence and Robotics for Society(深圳人工智能与机器人研究院) School of Science and Engineering(科学与工程学院) Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) ZJU-UIUC Institute(浙大-伊利诺伊大学联合学院) Zhejiang University(浙江大学) School of Management(管理学院) Huazhong University of Science and Technology(华中科技大学) Shenzhen Future Network of Intelligence Institute (FNii-Shenzhen)(深圳未来智能网络研究院(FNii-Shenzhen)) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Shenzhen Key Laboratory of Crowd Intelligence Empowered Low-Carbon Energy Network(深圳群智赋能低碳能源网络重点实验室) CSIJRI Joint Research Centre on Smart Energy Storage(CSIJRI联合智能储能研究中心)

AI总结 本文提出CSDP框架,通过关联序列差分隐私方法,在保持数据有用性的同时提升隐私保护效果。

Comments 11 pages, 5 figures. Published in 2025 34th International Conference on Computer Communications and Networks (ICCCN), IEEE, August 2025

Journal ref Proceedings of the 34th International Conference on Computer Communications and Networks (ICCCN 2025), IEEE, pp. 1-9, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06837 2025-11-25 cs.LG

Minimum Width of Deep Narrow Networks for Universal Approximation

深度窄网络的最小宽度用于通用逼近

Xiao-Song Yang, Qi Zhou, Xuan Zhou

机构 * Huazhong University of Science and Technology(华中科技大学)

AI总结 本文研究了深度窄网络的最小宽度以实现通用逼近能力,通过分析不同激活函数的特性,得出了新的宽度下界和上界结论。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19536 2025-11-25 cs.CV cs.AI cs.CL

FlowCut: Rethinking Redundancy via Information Flow for Efficient Vision-Language Models

FlowCut: 通过信息流重新思考冗余性以提高视觉-语言模型的效率

Jintao Tong, Wenwei Jin, Pengda Qin, Anqi Li, Yixiong Zou, Yuhong Li, Yuhua Li, Ruixuan Li

机构 * School of Computer Science and Technology, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院) Xiaohongshu Inc.(小红书公司) Shanghai Jiao Tong University(上海交通大学)

AI总结 FlowCut通过信息流视角改进视觉-语言模型的冗余识别,实现更高效的剪枝效果。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17199 2025-11-24 cs.CV

VLA-4D: Embedding 4D Awareness into Vision-Language-Action Models for SpatioTemporally Coherent Robotic Manipulation

VLA-4D: 将4D意识嵌入视觉-语言-动作模型中以实现时空一致的机器人操作

Hanyu Zhou, Chuanhao Ma, Gim Hee Lee

机构 * School of Computing, National University of Singapore(新加坡国立大学计算机学院) School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)

AI总结 VLA-4D通过4D意识增强视觉-语言-动作模型,实现时空一致的机器人操作,提升动作执行的空间平滑性和时间一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01588 2025-11-24 cs.LG cs.CV

Explore More, Learn Better: Parallel MLLM Embeddings under Mutual Information Minimization

探索更多,学习更好:基于互信息最小化的并行 MLLM 嵌入

Zhicheng Wang, Chen Ju, Xu Chen, Shuai Xiao, Jinsong Lan, Xiaoyong Zhu, Ying Chen, Zhiguo Cao

机构 * Zhejiang University(浙江大学) Alibaba Group(阿里巴巴集团) Huazhong University of Science and Technology(华中科技大学)

AI总结 本文提出基于互信息最小化的并行解耦框架,通过多条并行路径提升多模态嵌入质量,实现高效且鲁棒的嵌入空间。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18904 2025-11-21 cs.CV

TC-Light: Temporally Coherent Generative Rendering for Realistic World Transfer

TC-Light: 用于现实世界迁移的时序一致生成渲染

Yang Liu, Chuanchen Luo, Zimo Tang, Yingyan Li, Yuran Yang, Yuanyong Ning, Lue Fan, Junran Peng, Zhaoxiang Zhang

机构 * NLPR, MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) University of Chinese Academy of Sciences(中国科学院大学) Shandong University(山东大学) University of Science and Technology Beijing(北京科技大学) Tencent(腾讯) Huazhong University of Science and Technology(华中科技大学)

AI总结 TC-Light通过优化外观嵌入和唯一视频张量,实现高质量且高效的世界迁移渲染。

Comments Project Page: https://dekuliutesla.github.io/tclight/ Code: https://github.com/Linketic/TC-Light

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16163 2025-11-21 cs.CV

An Image Is Worth Ten Thousand Words: Verbose-Text Induction Attacks on VLMs

一张图像等于一万字:针对视觉语言模型的冗余文本诱导攻击

Zhi Luo, Zenghui Yuan, Wenqi Wei, Daizong Liu, Pan Zhou

机构 * Huazhong University of Science and Technology(华中科技大学) Fordham University(福特汉姆大学) Wuhan University(武汉大学)

AI总结 本文提出一种针对视觉语言模型的冗余文本诱导攻击,通过两阶段框架生成恶意图像以诱导模型生成冗长文本,提升攻击效果与可控性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15192 2025-11-21 cs.AI

As If We've Met Before: LLMs Exhibit Certainty in Recognizing Seen Files

似曾相识:LLM在识别已见过的文件时表现出确定性

Haodong Li, Jingqi Zhang, Xiao Cheng, Peihua Mai, Haoyu Wang, Yan Pang

机构 * Huazhong University of Science and Technology(华中科技大学) National University of Singapore(国立新加坡大学) Macquarie University(麦考瑞大学)

AI总结 COPYCHECK利用LLM的不确定性信号,通过双策略检测训练数据中的受版权内容,实现高准确率的版权检测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24473 2025-11-20 cs.CV cs.AI cs.CL cs.LG

Euclid's Gift: Enhancing Spatial Perception and Reasoning in Vision-Language Models via Geometric Surrogate Tasks

Shijie Lian, Changti Wu, Laurence Tianruo Yang, Hang Yuan, Bin Yu, Lei Zhang, Kai Chen

机构 * Huazhong University of Science and Technology(华中科技大学) Zhongguancun Academy(中关村学院) East China Normal University(华东师范大学) Zhengzhou University(郑州大学) Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14161 2025-11-20 cs.RO cs.CV

RoboTidy : A 3D Gaussian Splatting Household Tidying Benchmark for Embodied Navigation and Action

Xiaoquan Sun, Ruijian Zhang, Kang Pang, Bingchen Miao, Yuxiang Tan, Zhen Yang, Ming Li, Jiayu Chen

机构 * Huazhong University of Science and Technology(华中科技大学) The University of Hong Kong(香港大学) INFIFORCE Intelligent Technology Co., Ltd.(INFIFORCE智能技术有限公司) Zhejiang University(浙江大学) Guangming Lab, Shenzhen(深圳光明实验室)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.04483 2025-11-19 cs.AI cs.CL

GraphInstruct: Empowering Large Language Models with Graph Understanding and Reasoning Capability

Zihan Luo, Xiran Song, Hong Huang, Jianxun Lian, Chenhao Zhang, Jinqi Jiang, Xing Xie, Hai Jin

机构 * National Engineering Research Center for Big Data Technology and System(大数据技术与系统国家工程研究中心) System Lab, Cluster(系统实验室,集群) Grid Computing Lab, School of Computer Science(网格计算实验室,计算机科学学院) Huazhong University of Science and Technology(华中科技大学) School of Computer Science(计算机科学学院) Microsoft Research Asia, Beijing 100190, China(微软亚洲研究院,北京100190,中国)

Comments The article has been accepted by Frontiers of Computer Science (FCS), with the DOI: {10.1007/s11704-025-51382-0}

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18966 2025-11-18 cs.CL

DataGen: Unified Synthetic Dataset Generation via Large Language Models

Yue Huang, Siyuan Wu, Chujie Gao, Dongping Chen, Qihui Zhang, Yao Wan, Tianyi Zhou, Jianfeng Gao, Chaowei Xiao, Lichao Sun, Xiangliang Zhang

机构 * University of Notre Dame(诺丁汉大学) Huazhong University of Science and Technology(华中科技大学) MBZUAI University of Washington(华盛顿大学) Peking University(北京大学) University of Maryland, College Park(马里兰大学学院市分校) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) Microsoft Research(微软研究院) Lehigh University(莱斯大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12921 2025-11-18 cs.CV

Generative Photographic Control for Scene-Consistent Video Cinematic Editing

Huiqiang Sun, Liao Shen, Zhan Peng, Kun Wang, Size Wu, Yuhang Zang, Tianqi Liu, Zihao Huang, Xingyu Zeng, Zhiguo Cao, Wei Li, Chen Change Loy

机构 * School of AIA, Huazhong University of Science and Technology(华中科技大学人工智能学院) S-Lab, Nanyang Technological University(南洋理工大学S实验室) SenseTime Research(商汤科技研究院) Shanghai AI Laboratory(上海人工智能实验室)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12896 2025-11-18 cs.RO

Air-Chamber Based Soft Six-Axis Force/Torque Sensor for Human-Robot Interaction

Jun Huo, Hongge Ru, Bo Yang, Xingjian Chen, Xi Li, Jian Huang

机构 * Key Laboratory of the Ministry of Education for Image Processing and Intelligent Control(教育部图像处理与智能控制重点实验室) School of Artificial Intelligence and Automation(人工智能与自动化学院) Huazhong University of Science and Technology(华中科技大学)

Journal ref IEEE Transactions on Instrumentation and Measurement, vol. 73, pp. 1-12, 2024, Art no. 9501612,

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10390 2025-11-18 cs.CV cs.AI

MonkeyOCR v1.5 Technical Report: Unlocking Robust Document Parsing for Complex Patterns

Jiarui Zhang, Yuliang Liu, Zijun Wu, Guosheng Pang, Zhili Ye, Yupei Zhong, Junteng Ma, Tao Wei, Haiyang Xu, Weikai Chen, Zeen Wang, Qiangjun Ji, Fanxi Zhou, Qi Zhang, Yuanrui Hu, Jiahao Liu, Zhang Li, Ziyang Zhang, Qiang Liu, Xiang Bai

机构 * KingSoft Office Zhuiguang AI Lab(金山办公紫光人工智能实验室) Huazhong University of Science and Technology(华中科技大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06543 2025-11-17 cs.CV

MILD: Multi-Layer Diffusion Strategy for Complex and Precise Multi-IP Aware Human Erasing

Jinghan Yu, Junhao Xiao, Zhiyuan Ma, Yue Ma, Kaiqi Liu, Yuhan Wang, Daizong Liu, Xianghao Meng, Jianjun Li

机构 * Huazhong University of Science and Technology(华中科技大学) Tsinghua University(清华大学) Hong Kong University of Science and Technology(香港科技大学) Nanjing University(南京大学) Wuhan University(武汉大学) Central China Normal University(中央财经大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25238 2025-11-14 cs.CV

VADB: A Large-Scale Video Aesthetic Database with Professional and Multi-Dimensional Annotations

Qianqian Qiao, DanDan Zheng, Yihang Bo, Bao Peng, Heng Huang, Longteng Jiang, Huaye Wang, Jingdong Chen, Jun Zhou, Xin Jin

机构 * Nanjing University(南京大学) Huazhong University of Science and Technology(华中科技大学) Beijing Film Academy(北京电影学院) University of Science and Technology of China(中国科学技术大学) Beijing Electronic Science and Technology Institute(北京电子科技学院) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Beijing Institute for General Artificial Intelligence(北京通用人工智能研究院)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06817 2025-11-13 cs.CV cs.AI

TiS-TSL: Image-Label Supervised Surgical Video Stereo Matching via Time-Switchable Teacher-Student Learning

Rui Wang, Ying Zhou, Hao Wang, Wenwei Zhang, Qiang Li, Zhiwei Wang

机构 * Wuhan National Laboratory for Optoelectronics, Huazhong University of Science and Technology(华中科技大学光电实验室) Wuhan United Imaging Surgical Co., Ltd.(武汉联合影像手术有限公司)

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25822 2025-11-13 cs.RO

Act to See, See to Act: Diffusion-Driven Perception-Action Interplay for Adaptive Policies

Jing Wang, Weiting Peng, Jing Tang, Zeyu Gong, Xihua Wang, Bo Tao, Li Cheng

机构 * University of Alberta(阿尔伯塔大学) Huazhong University of Science and Technology(华中科技大学)

Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15526 2025-11-13 eess.IV cs.CV

Multi-scale Cascaded Foundation Model for Whole-body Organs-at-risk Segmentation

Rui Hao, Dayu Tan, Qiankun Li, Chunhou Zheng, Weimin Zhong, Zhigang Zeng

机构 * School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(人工智能与自动化学院,华中科技大学) Institute of Artificial Intelligence, Huazhong University of Science and Technology(人工智能研究院,华中科技大学) Hubei Key Laboratory of Brain-Inspired Intelligent Systems, Huazhong University of Science and Technology(湖北省脑启发智能系统重点实验室,华中科技大学) Key Laboratory of Image Processing and Intelligent Control (Huazhong University of Science and Technology), Ministry of Education(图像处理与智能控制重点实验室(华中科技大学),教育部) Key Laboratory of Intelligent Computing and Signal Processing, Ministry of Education, Anhui University(智能计算与信号处理重点实验室(安徽大学),教育部) College of Computing and Data Science (CCDS), Nanyang Technological University(计算与数据科学学院(CCDS),南洋理工大学) East China University of Science and Technology(东华大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22340 2025-11-12 cs.AI cs.CL cs.CV cs.LG

DynaSolidGeo: A Dynamic Benchmark for Genuine Spatial Mathematical Reasoning of VLMs in Solid Geometry

Changti Wu, Shijie Lian, Zihao Liu, Lei Zhang, Laurence Tianruo Yang, Kai Chen

机构 * East China Normal University(华东师范大学) Zhongguancun Academy(中关村学院) Huazhong University of Science and Technology(华中科技大学) Peking University(北京大学) Zhengzhou University(郑州大学) Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究院)

Comments The code and dataset are available at \href{https://zgca-ai4edu.github.io/DynaSolidGeo/}{DynaSolidGeo}

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14686 2025-11-12 cs.CV

From Semantics, Scene to Instance-awareness: Distilling Foundation Model for Grounded Open-vocabulary Situation Recognition

Chen Cai, Tianyi Liu, Jianjun Gao, Wenyang Liu, Kejun Wu, Ruoyu Wang, Yi Wang, Soo Chin Liew

机构 * National University of Singapore(新加坡国立大学) Nanyang Technological University(南洋理工大学) Huazhong University of Science and Technology(华中科技大学) The Hong Kong Polytechnic University(香港理工大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06821 2025-11-11 math.GN cs.LG

Dimensionality reduction and width of deep neural networks based on topological degree theory

Xiao-Song Yang

机构 * School of Mathematics and Statistics, Huazhong University of Science and Technology(数学与统计学学院,华中科技大学) Hubei Key Laboratory of Engineering Modeling and Scientific Computing , Huazhong University of Science and Technology(工程建模与科学计算湖北省重点实验室,华中科技大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11613 2025-11-11 cs.CV

High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network

Feng Zhang, Haoyou Deng, Zhiqiang Li, Lida Li, Bin Xu, Qingbo Lu, Zisheng Cao, Minchen Wei, Changxin Gao, Nong Sang, Xiang Bai

机构 * National Key Laboratory of Multispectral Information Intelligent Processing Technology, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(国家多谱段信息智能处理技术重点实验室,人工智能与自动化学院,华中科技大学) DJI Technology Co., Ltd.(大疆技术创新有限公司) Color, Imaging, and Illumination Laboratory, The Hong Kong Polytechnic University(色彩、成像与照明实验室,香港理工大学) School of Software Engineering, Huazhong University of Science and Technology(软件工程学院,华中科技大学)

Comments accepted by TPAMI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04263 2025-07-08 cs.RO

SRefiner: Soft-Braid Attention for Multi-Agent Trajectory Refinement

Liwen Xiao, Zhiyu Pan, Zhicheng Wang, Zhiguo Cao, Wei Li

机构 * School of AIA, Huazhong University of Science and Technology(华中科技大学人工智能学院) S-Lab, Nanyang Technological University(南洋理工大学S实验室)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04118 2025-07-08 cs.CV

PromptSR: Cascade Prompting for Lightweight Image Super-Resolution

Wenyang Liu, Chen Cai, Jianjun Gao, Kejun Wu, Yi Wang, Kim-Hui Yap, Lap-Pui Chau

机构 * School of Electrical and Electronics Engineering, Nanyang Technological University, Singapore(南洋理工大学电子与电气工程学院) School of Electronic Information and Communications, Huazhong University of Science and Technology, Wuhan 430074, China(华中科技大学电子信息与通信学院) Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University, Hong Kong(香港理工大学电子与电气工程系)

Comments Accepted in TMM

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03893 2025-07-08 cs.CV cs.AI

Hierarchical Semantic-Visual Fusion of Visible and Near-infrared Images for Long-range Haze Removal

Yi Li, Xiaoxiong Wang, Jiawei Wang, Yi Chang, Kai Cao, Luxin Yan

机构 * National Key Laboratory of Multispectral Information Intelligent Processing Technology, School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(多光谱信息智能处理技术国家实验室,人工智能与自动化学院,华中科技大学) State Key Laboratory of Dynamic Optical Imaging and Measurement(动态光学成像与测量国家重点实验室)

Comments This work has been accepted by IEEE Transactions on Multimedia for publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10943 2025-07-08 cs.CV

Rethinking Detecting Salient and Camouflaged Objects in Unconstrained Scenes

Zhangjun Zhou, Yiping Li, Chunlin Zhong, Jianuo Huang, Jialun Pei, Hua Li, He Tang

机构 * School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院) School of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程学院) School of Computer Science and Technology, Hainan University(海南大学计算机科学与技术学院)

Comments 17 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏