机构
*
The University of Hong Kong(香港大学)
;
Shenyang Institute of Automation, Chinese Academy of Sciences(中国科学院沈阳自动化研究所)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
Scaling Parallel Sequence Models to Foundation-Scale Vision Encoders
将并行序列模型扩展到基础规模的视觉编码器
Yitong Jiang, Hongjun Wang, Collin McCarthy, Hanrong Ye, David Wehr, Xinhao Li, Qi Dou, Tianfan Xue, Ka Chun Cheung, Simon See, Wonmin Byeon, Ke Chen, Kai Han, Jinwei Gu, Hongxu Yin, Pavlo Molchanov, Jan Kautz, Sifei Liu
机构
*
NVIDIA
;
The Chinese University of Hong Kong(香港中文大学)
;
The University of Hong Kong(香港大学)
;
University of California, San Diego(加州大学圣地亚哥分校)
Fangyuan Wang, Ziyuan Wang, Guorui Pei, Mengshi Zhang, Canxi Liang, Jun Hu, Zhongxuan Li, Jinsong Wu, Ning Han, Zeqing Zhang, Jiaming Qi, Hongmin Wu, Shiyao Zhang, Pai Zheng, Jia Pan, David Navarro-Alarcon, Sichao Liu, Peng Zhou
机构
*
Department of Mechanical Engineering, The Hong Kong Polytechnic University(香港理工大学机械工程系)
;
Department of Mechanical Engineering and Automation, Harbin Institute of Technology(哈尔滨工业大学机械工程与自动化系)
;
School of Advanced Engineering, Great Bay University(大湾大学先进工程学院)
;
College of Robotics Science and Engineering, Taiyuan University of Technology(太原科技大学机器人科学与工程学院)
;
School of Data Science, City University of Hong Kong (Dongguan)(香港城市大学(东莞)数据科学学院)
;
Department of Mechatronic Engineering, Guangdong Polytechnic Normal University(广东 polytechnic 正常大学机电工程系)
;
School of Computing and Data Science, The University of Hong Kong(香港大学计算与数据科学学院)
;
School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)
;
College of Mechanical and Electrical Engineering, Northeast Forestry University(东北林业大学机械与电气工程学院)
;
Greater Bay Area National Center of Technology Innovation(粤港澳大湾区国家技术创新中心)
;
Department of Industrial and Systems Engineering, The Hong Kong Polytechnic University(香港理工大学工业与系统工程系)
机构
*
Department of Mathematics Faculty of Science(科学学院数学系)
;
The University of Hong Kong Hong Kong SAR(香港大学香港特别行政区)
;
School of Computing and Mathematical Sciences Faculty of Engineering and Science(工程与科学学院计算与数学科学系)
;
University of Greenwich United Kingdom(格林威治大学英国)
机构
*
Institute of Trustworthy Embodied AI (TEAI)(可信具身人工智能研究院)
;
Shanghai Key Laboratory of Multimodal Embodied AI(上海多模态具身人工智能重点实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
OpenDriveLab, The University of Hong Kong(OpenDrive实验室,香港大学)
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
看、规划、回退:面向鲁棒机器人操作的进度感知视觉-语言-动作模型
Tingjun Dai, Mingfei Han, Tingwen Du, Zhiheng Liu, Zihao Zhang, Zhihui Li, Salman Khan, Jun Yu, Xiaojun Chang
机构
*
School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学)
;
University of Technology Sydney(新南威尔士大学)
;
Department of Computer Vision, Mohamed Bin Zayed University of Artificial Intelligence(人工智能与计算机视觉系,Mohamed Bin Zayed人工智能大学)
;
The University of Hong Kong(香港大学)
;
Institute of AI for Industry, Chinese Academy of Sciences(产业人工智能研究所,中国科学院)
;
School of Intelligent Science and Engineering, Harbin Institute of Technology (Shenzhen)(智能科学与工程学院,哈尔滨工业大学(深圳))
HuMam: Humanoid Motion Control via End-to-End Deep Reinforcement Learning with Mamba
HuMam: 基于Mamba的端到端深度强化学习人形机器人运动控制
Yinuo Wang, Yuanyang Qi, Jinzhao Zhou, Pengxiang Meng, Xiaowen Tao
机构
*
College of Graduate and Professional Studies, Trine University(特灵大学研究生与专业研究学院)
;
Department of Civil Engineering, University of Hong Kong(香港大学土木工程系)
;
Faculty of Engineering and Information Technology, University of Technology Sydney(悉尼大学工程与信息技术学院)
;
National Key Laboratory of Automotive Chassis Integration and Bionics, Jilin University(吉林大学汽车底盘集成与生物仿生国家重点实验室)
;
School of Computer Science and Statistics, Trinity College Dublin(都柏林信任学院计算机科学与统计学系)
Journal ref2026 IEEE International Conference on Cybernetics and Intelligent Systems (CIS) and IEEE International Conference on Robotics, Automation and Mechatronics (RAM) (CIS-RAM 2026)
DAG-Plan: Generating Directed Acyclic Dependency Graphs for Dual-Arm Cooperative Planning
DAG-Plan:生成有向无环依赖图用于双臂协作规划
Zeyu Gao, Yao Mu, Jinye Qu, Mengkang Hu, Shijia Peng, Chengkai Hou, Lingyue Guo, Ping Luo, Shanghang Zhang, Yanfeng Lu
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences (CASIA)(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院(CASIA))
;
School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS)(中国科学院大学人工智能学院)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,北京大学计算机科学学院)
;
Department of Computer Science, The University of Hong Kong(香港大学计算机科学系)
;
OpenGVLab, Shanghai AI Laboratory(上海人工智能实验室,OpenGVLab)
机构
*
School of Mathematical Sciences, Peking University(北京大学数学科学学院)
;
Department of Computer Science, University of California, Los Angeles(加州大学洛杉矶分校计算机科学系)
;
School of Computing and Data Science, the University of Hong Kong(香港大学计算科学与数据科学学院)
;
Bytedance Inc(字节跳动公司)
AI总结
提出 AdaGrad++ 和 Adam++ 两种简单无参数自适应梯度方法,在无需预设学习率的情况下实现与 AdaGrad 和 Adam 相当的收敛保证。