ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training
ALOE: 面向视觉-语言-动作模型后训练的动作级离策略评估
Rushuai Yang, Hecheng Wang, Zhichao Wu, Chiming Liu, Xiaohan Yan, Xuan Du, Shuoyu Yue, Chuheng Zhang, Yunlong Wang, Yongcheng Liu, Lizhe Qi, Yi Chen, Wei Shan, Maoqing Yao
机构
*
AgiBot
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Fudan University(复旦大学)
;
Nanjing University(南京大学)
;
Independent Researcher(独立研究者)
机构
*
The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳))
;
School of Data Science, School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen, China(数据科学学院、人工智能学院、香港中文大学(深圳))
机构
*
School of Optics and Photonics, Beijing Institute of Technology(光学与光子学学院,北京理工大学)
;
School of Optoelectronic Engineering, Changchun University of Science and Technology(光电工程学院,长春理工大学)
Can Context Bridge the Reality Gap? Sim-to-Real Transfer of Context-Aware Policies
上下文能否弥合现实差距?情境感知策略的模拟到现实迁移
Marco Iannotta, Yuxuan Yang, Johannes A. Stork, Erik Schaffernicht, Todor Stoyanov
机构
*
AASS Research Centre, Örebro University(奥雷布罗大学AASS研究中心)
;
Technology Transfer Center Kitzingen, Technical University of Applied Sciences Würzburg-Schweinfurt(基廷根技术转移中心,沃尔夫斯堡-施维林应用技术大学)
On Motion Blur and Deblurring in Visual Place Recognition
视觉场所识别中的运动模糊与去模糊
Timur Ismagilov, Bruno Ferrarini, Michael Milford, Tan Viet Tuyen Nguyen, SD Ramchurn, Shoaib Ehsan
机构
*
School of Electronics and Computer Science, University of Southampton(苏塞克斯大学电子与计算机科学学院)
;
MyWay srl(MyWay公司)
;
QUT Centre for Robotics, School of Electrical Engineering and Robotics(昆士兰大学机器人中心,电气工程与机器人学院)
;
School of Computer Science and Electronic Engineering, University of Essex(埃塞克斯大学计算机科学与电子工程学院)
机构
*
Qwen Business Unit of Alibaba(阿里巴巴通义千问业务部)
;
ShanghaiTech University(上海科技大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Institute of Computing Technology(中国科学院计算技术研究所)
;
Southeast University(东南大学)
Toward Universal Skeleton-Based Action Recognition across Heterogeneous Skeletons and Open Vocabularies
迈向通用的基于骨骼的动作识别
Jidong Kuang, Hongsong Wang, Jie Gui, Yuan Yan Tang, James Tin-Yau Kwok
机构
*
School of Cyber Science and Engineering, Southeast University, Nanjing, China(东南大学计算机科学与工程学院,南京,中国)
;
School of Computer Science and Engineering, Southeast University, Nanjing, China(东南大学计算机科学与工程学院,南京,中国)
Comments10 pages, 4 figures; accepted at the CVPR 2026 Workshop on Video Generative Models: Benchmarks and Evaluation (VGBE). Updated to the complete camera-ready version
Journal refProceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), 2026
PerFACT: Motion Policy with LLM-Powered Dataset Synthesis and Fusion Action-Chunking Transformers
PerFACT: 基于LLM驱动的数据集合成与融合动作分块变换器的运动策略
Davood Soleymanzadeh, Xiao Liang, Minghui Zheng
机构
*
J. Mike Walker ’66 Department of Mechanical Engineering, Texas A&M University(J. Mike Walker ’66机械工程系,德克萨斯A&M大学)
;
Zachry Department of Civil and Environmental Engineering, Texas A&M University(Zachry土木与环境工程系,德克萨斯A&M大学)
Data-driven registration and modeling of brain deformation for image-guided neurosurgery
数据驱动的图像配准与变形建模在图像引导神经外科中的应用:系统综述
Tiago Assis, Colin P. Galvin, Joshua P. Castillo, Nazim Haouchine, Marta Kersten-Oertel, Zeyu Gao, Mireia Crispin-Ortuzar, Stephen J. Price, Thomas Santarius, Yangming Ou, Sarah Frisken, Nuno C. Garcia, Alexandra J. Golby, Reuben Dorent, Ines P. Machado
机构
*
LASIGE, Faculty of Sciences, University of Lisbon(里斯本大学科学学院LASIGE)
;
Department of Neurosurgery and Department of Radiology, Brigham and Women's Hospital, Harvard Medical School(哈佛医学院布里洛妇女医院神经外科与放射科)
;
Gina Cody School of Engineering and Computer Science, Concordia University(康科迪亚大学工程与计算机科学学院)
;
Cancer Research UK Cambridge Centre, University of Cambridge(剑桥大学癌症研究英国中心)
;
Department of Oncology, University of Cambridge(剑桥大学肿瘤科)
;
Department of Clinical Neurosciences, University of Cambridge(剑桥大学临床神经科学系)
;
Computational Health Informatics Program (CHIP) and Department of Radiology, Boston Children's Hospital, Harvard Medical School(哈佛医学院波士顿儿童医院计算健康信息学计划与放射科)
;
Sorbonne Université, Institut du Cerveau - Paris Brain Institute - ICM(索邦大学巴黎脑研究所-ICM)
Journal refAssis, T., et al. (2026). Data-driven registration and modeling of brain deformation for image-guided neurosurgery. Medical Image Analysis, 114, 104217
GIFGuard: Proactive Forensics against Deepfakes in Facial GIFs via Spatiotemporal Watermarking
GIFGuard:通过时空水印实现面向面部GIF的主动取证对抗深度伪造
Shupeng Che, Zhiqing Guo, Changtao Miao, Dan Ma, Gaobo Yang
机构
*
School of Computer Science and Technology, Xinjiang University(新疆大学计算机科学与技术学院)
;
College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院)
Yan Zhou, Suncheng Xiang, Zhen Huang, Yue Ouyang, Yingqiu Li, Zehua Wang
机构
*
School of Mathematics and Statistics, Changsha University of Science and Technology(数学与统计学学院,长沙理工大学)
;
School of Biomedical Engineering, Shanghai Jiao Tong University(生物医学工程学院,上海交通大学)
;
Shanghai Chest Hospital, Shanghai Jiao Tong University School of Medicine(上海胸科医院,上海交通大学医学院)
Legible and Intuitive Multi-modal Robot State and Intent Communication Validated in Online and Real-world Studies
可读且直观的多模态机器人状态与意图通信:在线和真实世界研究验证
Tim Schreiter, Jens V. Rüppel, Andrey Rudenko, Martin Magnusson, Achim J. Lilienthal
机构
*
Chair of Perception for Intelligent Systems, Munich Institute of Robotics and Machine Intelligence (MIRMI), Technical University of Munich (TUM)(慕尼黑工业大学慕尼黑机器人与机器智能研究所智能系统感知教席)
;
Centre for Applied Autonomous Sensor Systems (AASS), Örebro University(厄勒布鲁大学应用自主传感器系统中心)
;
Robotics Institute Germany (RIG)(德国机器人研究所)