arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Huazhong University of Science and Technology(华中科技大学)

2026-04-30 至 2026-04-30 共收录 6
2604.26917 2026-04-30 cs.CV

AnimateAnyMesh++: A Flexible 4D Foundation Model for High-Fidelity Text-Driven Mesh Animation

AnimateAnyMesh++: 一种灵活的4D基础模型用于高质量文本驱动的网格动画

Zijie Wu, Chaohui Yu, Fan Wang, Xiang Bai

机构 * School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院) DAMO Academy, Alibaba Group(阿里巴巴达摩院) Hupan Lab, Hangzhou, China(湖畔实验室) School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)

AI总结 本文提出AnimateAnyMesh++,通过扩展数据集、改进架构和生成能力,实现高质量文本驱动的网格动画,提升了轨迹重建和几何保真度。

Comments 14 pages, TPAMI submission, code url: https://github.com/JarrentWu1031/AnimateAnyMesh-pp

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26806 2026-04-30 cs.CV cs.AI

ViCrop-Det: Spatial Attention Entropy Guided Cropping for Training-Free Small-Object Detection

ViCrop-Det: 基于空间注意力熵引导的裁剪用于无训练小目标检测

Hui Wang, Hongze Li, Wei Chen, Xiaojin Zhang

机构 * School of Computer Science and Technology, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院) School of Software Engineering, Huazhong University of Science and Technology(华中科技大学软件工程学院)

AI总结 ViCrop-Det通过空间注意力熵引导裁剪,无需训练即可提升小目标检测性能,实验表明在VisDrone和DOTA-v1.5数据集上,其性能提升显著且计算开销可控。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26707 2026-04-30 cs.CV cs.LG

CurEvo: Curriculum-Guided Self-Evolution for Video Understanding

CurEvo: 基于课程引导的视频理解自进化

Guiyi Zeng, Junqing Yu, Yi-Ping Phoebe Chen, Xu Chen, Wei Yang, Zikai Song

机构 * Huazhong University of Science and Technology(华中科技大学) La Trobe University(拉特罗布大学) Beijing Institute of Computer Technology and Applications(北京计算机技术与应用研究院)

AI总结 CurEvo通过引入课程学习提升视频理解的自进化过程,实现更结构化的模型改进,验证了课程引导自进化在视频理解中的有效性。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26353 2026-04-30 cs.CV

GateMOT: Q-Gated Attention for Dense Object Tracking

Mingjin Lv, Zelin Liu, Feifei Shao, Yi-Ping Phoebe Chen, Junqing Yu, Wei Yang, Zikai Song

机构 * Huazhong University of Science and Technology, Wuhan, China(华中科技大学,武汉,中国) Zhejiang University, Hangzhou, China(浙江大学,杭州,中国) La Trobe University, Melbourne, Australia(拉筹伯大学,墨尔本,澳大利亚)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26252 2026-04-30 cs.CV

OmniTrend: Content-Context Modeling for Scalable Social Popularity Prediction

OmniTrend: 内容-上下文建模用于可扩展的社会流行度预测

Liliang Ye, Guiyi Zeng, Yunyao Zhang, Yi-Ping Phoebe Chen, Junqing Yu, Zikai Song

机构 * Huazhong University of Science and Technology(华中科技大学) La Trobe University(拉特罗布大学)

AI总结 本文提出OmniTrend框架,通过分离内容吸引力与上下文曝光,提升跨平台流行度预测的准确性和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07080 2026-04-30 cs.RO cs.LG

VLN-Cache: Enabling Token Caching for VLN Models with Visual/Semantic Dynamics Awareness

VLN-Cache: 通过视觉/语义动态感知实现VLN模型的令牌缓存

Zihao Zheng, Zhihao Mao, Xingyue Zhou, Jiayu Chen, Maoliang Li, Xinhao Sun, Hailong Zou, Zhaobo Zhang, Xuanzhe Liu, Donggang Cao, Hong Mei, Xiang Chen

机构 * School of Computer Science, Peking University(北京大学计算机科学系) School of Computer Science, China University of Geosciences (Wuhan)(中国地质大学(武汉)计算机科学系) School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院) School of Electronics Engineering and Computer Science, Peking University(北京大学电子工程与计算机科学系)

AI总结 本文提出VLN-Cache框架,通过视觉动态感知和语义动态感知解决VLN模型中因视角和语义变化导致的缓存失效问题,实验显示在R2R-CE模拟基准上实现1.52倍的速度提升。

详情

展开后加载摘要…

URL PDF HTML 收藏