Comments8 pages, 6 figures. Accepted manuscript. Published in the 2025 IEEE-RAS 24th International Conference on Humanoid Robots (Humanoids), pp. 1233-1240
Journal ref2025 IEEE-RAS 24th International Conference on Humanoid Robots (Humanoids), pp. 1233-1240 (2025)
Evolving Cache Schedules for Fast Diffusion Policy Inference
用于快速扩散策略推理的演进缓存调度
Siying Wang, Kangye Ji, Di Wang, Fei Cheng
机构
*
School of Telecommunications Engineering, Xidian University(西安电子科技大学通信工程学院)
;
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
机构
*
GigaAI(字节跳动人工智能实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
University of Leeds(利兹大学)
;
Cornell University(康奈尔大学)
;
Tsinghua University(清华大学)
;
Nanjing University of Science and Technology(南京理工大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
FAWTD(一汽技术开发部)
CommentsWe are withdrawing this manuscript because our current analysis has revealed serious analytical issues that render the main conclusions invalid. We intend to thoroughly revise the methodology and submit a corrected version in the future
Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning
不要愚弄我两次:通过经验驱动推理在野外适应逆境
Navin Sriram Ravie, Andrew Jong, Krrish Jain, John Liu, Omar Alama, Bijo Sebastian, Sebastian Scherer
机构
*
Department of Engineering Design, Indian Institute of Technology, Madras(印度理工学院工程设计系,马德拉斯)
;
Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所)
Interactive Medical-SAM2 GUI: A Napari-based semi-automatic annotation tool for medical images
交互式医学-SAM2 GUI:基于Napari的半自动标注工具用于医学图像
Woojae Hong, Jong Ha Hwang, Jiyong Chung, Joongyeon Choi, Hyunggun Kim, Yong Hwy Kim
机构
*
Department of Biomechatronic Engineering, Sungkyunkwan University(生物机电工程系,全州大学)
;
Department of Neurosurgery, Seoul National University Hospital, Seoul National University College of Medicine(神经外科,首尔国立大学医院,首尔国立大学医学院)
KineBench: Benchmarking Embodied World Models via IDM-Free Kinematic Grounding
KineBench:通过无逆动力学模型的运动学基础对具身世界模型进行基准测试
Zeyu Liu, Zhangzhe Zhu, Yang Zhang, Chenyou Fan, Chenjia Bai, Xuelong Li
机构
*
Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究院)
;
National University of Singapore(新加坡国立大学)
;
Fudan University(复旦大学)
;
Tsinghua University(清华大学)
;
Shenzhen Research Institute of Northwestern Polytechnical University(西北工业大学深圳研究院)
机构
*
Graduate School of Informatics, Kyoto University, Kyoto, Japan(京都大学信息学研究生院)
;
Research Organization of Science and Technology, Ritsumeikan University, Shiga, Japan(立命馆大学科学技术研究机构)
HMVLA: Hyperbolic Multimodal Fusion for Vision-Language-Action Models
HMVLA:超几何多模态融合用于视觉-语言-动作模型
Kun Wang, Xiao Feng, Mingcheng Qu, Tonghua Su
机构
*
Harbin Institute of Technology(哈尔滨工业大学)
;
Chongqing Research Institute of HIT(重庆HIT研究 institute)
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室(深圳))
专题命中
机器人基础模型
:robotic(abstract);分类 cs.RO、cs.LG
AI总结
HMVLA通过在双曲空间中融合视觉、语言和动作信息,提升多模态语义对齐的效率和准确性。
Comments5 pages,5 figures,ICASSP
Journal refICASSP 2026 - 2026 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)