Skeleton-to-Image Encoding: Enabling Skeleton Representation Learning via Vision-Pretrained Models
骨架到图像编码:通过视觉预训练模型实现骨架表示学习
Siyuan Yang, Jun Liu, Hao Cheng, Chong Wang, Shijian Lu, Hedvig Kjellstrom, Weisi Lin, Alex C. Kot
机构
*
KTH Royal Institute of Technology(皇家理工学院)
;
Lancaster University(兰卡斯特大学)
;
Hebei University of Technology(河北工业大学)
;
Nanyang Technological University(南洋理工大学)
;
Shenzhen MSU-BIT University(深圳MSU-BIT大学)
;
VinUniversity(文大学)
Comments17 pages, 14 figures, accepted to Computer Vision and Pattern Recognition Conference (CVPR) Workshops 2026. 5th MMFM Workshop: What is Next in Multimodal Foundation Models?
Journal refIn Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (pp. 7415-7424) 2026
Is Self-Pretraining really useful to improve diagnosis in medical Time Series?
自预训练(SPT)真的有助于改进医疗时间序列的诊断吗?
Omar Coser, Antonio Orvieto, Paolo Soda, Loredana Zollo
机构
*
Università Campus Bio-Medico di Roma(罗马生物医学大学校园大学)
;
Umeå University(于默奥大学)
;
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
ELLIS Institute Tübingen(埃利斯研究所蒂宾根分所)
Commentsv5: Fixed PDF rendering compatibility issue affecting Apple PDFKit (macOS Preview/iOS PDF viewer). No changes to technical content compared to v4
机构
*
School of Information Science and Technology, Yunnan Normal University(云南师范大学信息科学与技术学院)
;
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
;
College of Computer Science, Beijing University of Technology(北京工业大学计算机学院)
ToDMA: Large Model-Driven Massive Token Communications for Semantic Multiple Access
ToDMA:用于语义多址接入的大模型驱动海量令牌通信
Li Qiao, Mahdi Boloursaz Mashhadi, Zhen Gao, Robert Schober, Deniz Gündüz
机构
*
The University of Hong Kong(香港大学)
;
Beijing Institute of Technology(北京理工大学)
;
University of Surrey(Surrey大学)
;
Friedrich-Alexander-University Erlangen-Nurnberg(埃森哲-亚琛工业大学)
;
Imperial College London(伦敦帝国理工学院)
机构
*
School of Automation, Beijing Institute of Technology(自动化学院,北京理工大学)
;
School of Computer Science, Wuhan University(计算机学院,武汉大学)
;
Great Wall Motor(长城汽车)
;
School of Information and Electronic Engineering, Zhejiang University of Science and Technology(信息电子工程学院,浙江理工大学)
CommentsThis paper is an extended version of the authors' work previously presented at the ICRA conference. To appear in IEEE Transactions on Circuits and Systems for Video Technology. DOl: 10.1109/TCSVT.2026.3701706
SpatialFly: Implicit 3D Prior-Guided Visual Reparameterization for Continuous UAV Vision-and-Language Navigation
SpatialFly:基于几何的表示对齐用于城市环境中无人机视觉与语言导航
Wen Jiang, Kangyao Huang, Li Wang, Wang Xu, Wei Fan, Jinyuan Liu, Shaoyu Liu, Hanfang Liang, Hongwei Duan, Bin Xu, Xiangyang Ji, Huaping Liu
机构
*
School of Mechanical Engineering, Beijing Institute of Technology(北京理工大学机械工程学院)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Tsinghua University(清华大学)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
School of Artificial Intelligence, Xidian University(西安电子科技大学人工智能学院)
;
Jianghan University(江汉大学)
Gated Relational Alignment via Confidence-based Distillation for Efficient VLMs
基于置信度蒸馏的门控关系对齐用于高效视觉语言模型
Yanlong Chen, Amirhossein Habibian, Luca Benini, Yawei Li
机构
*
Department of Information Technology(信息科技系)
;
Electrical Engineering, ETH Zurich, Zurich, Switzerland(电气工程,苏黎世联邦理工学院,苏黎世,瑞士)
;
Qualcomm AI Research, Amsterdam, the Netherlands(高通人工智能研究,阿姆斯特丹,荷兰)
;
Department of Electrical, Electronic and Information Engineering(电气、电子与信息工程系)
;
University of Bologna, Bologna, Italy(博洛尼亚大学,博洛尼亚,意大利)
;
School of Electrical and Electronic Engineering(电气与电子工程学院)
机构
*
Shanghai Jiao Tong University China
;
Hohai University China
;
Singapore Management University Singapore
;
Imperial College London United Kingdom
;
East China Normal University \& Shanghai Innovation Institute China
;
Chongqing University China
;
Shanghai Jiao Tong University
;
Hohai University
;
Singapore Management University
;
Imperial College London
;
East China Normal University \& Shanghai Innovation Institute
;
Chongqing University
Ming Wen, Yuxuan Liu, Kun Yang, Yunhao Feng, Zhuoer Xu, Yuhao Sun, Shiwen Cui, Xiang Zheng, Yi Liu, Xingjun Ma, Yu-Gang Jiang
机构
*
Institute of Trustworthy Embodied AI(可信具身人工智能研究院)
;
Fudan University(复旦大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Ant Group(蚂蚁集团)
;
Zhejiang University(浙江大学)
;
City University of Hong Kong(香港城市大学)