Vision-Language-Policy Model for Dynamic Robot Task Planning
用于动态机器人任务规划的视觉-语言-策略模型
Jin Wang, Kim Tien Ly, Jacques Cloete, Jin Jin, Nikos Tsagarakis, Ioannis Havoutis
机构
*
Dynamic Robot Systems Group, Oxford Robotics Institute, University of Oxford(牛津大学机器人研究所动态机器人系统组)
;
Humanoids and Human-Centered Mechatronics (HHCM), Istituto Italiano di Tecnologia(意大利技术研究所人形与以人为中心的机电系统)
Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction
球面-GOF:面向3D场景重建的几何感知全景高斯不透明场
Zhe Yang, Guoqiang Zhao, Sheng Wu, Kai Luo, Kailun Yang
机构
*
School of Artificial Intelligence and Robotics, Hunan University(人工智能与机器人学院,湖南大学)
;
National Engineering Research Center of Robot Visual Perception and Control Technology, Hunan University(机器人视觉感知与控制技术国家工程研究中心,湖南大学)
You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector
你有一张金票:用单一噪声向量改进生成式机器人策略
Omkar Patil, Ondrej Biza, Thomas Weng, Karl Schmeckpeper, Wil Thomason, Xiaohan Zhang, Kausik Sivakumar, Robin Walters, Nakul Gopalan, Sebastian Castro, Stephen Hart, Eric Rosen
机构
*
Robotics and AI Institute(机器人与人工智能研究所)
;
Arizona State University(亚利桑那州立大学)
;
Northeastern University(东北大学)
Simon Sagmeister, Marcel Weinmann, Phillip Pitschi, Markus Lienkamp
机构
*
Technical University of Munich, Germany(慕尼黑技术大学)
;
School of Engineering & Design, Department of Mobility Systems Engineering, Institute of Automotive Technology(工程与设计学院,移动系统工程系,汽车技术研究所)
;
School of Engineering & Design, Department of Engineering Physics and Computation, Institute of Automatic Control(工程与设计学院,工程物理与计算系,自动控制研究所)
CommentsWe have further refined the benchmark construction and experimental presentation to improve clarity and consistency. The revised version includes updated task design, food-resource data, and evaluation details to better align the benchmark with the intended food resource referral setting. These changes provide a more precise presentation of the experimental findings
JailWAM: Jailbreaking World Action Models in Robot Control
JailWAM: 在机器人控制中对World Action Models进行劫持
Hanqing Liu, Songping Wang, Jiahuan Long, Jiacheng Hou, Jialiang Sun, Chao Li, Yang Yang, Wei Peng, Xu Liu, Tingsong Jiang, Yao Mu, Wen Yao
机构
*
MoE key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院教育部重点实验室)
;
PR Lab, Nanjing University(南京大学PR实验室)
;
Defense Innovation Institute, Chinese Academy of Military Science(军事科学院国防创新研究院)
Breaking Déjà Vu: Independent Auditing of Visual Place Recognition through Vision-Language Reasoning
打破似曾相识:通过视觉语言推理对视觉场所识别进行独立审计
Sania Waheed, Michael Milford, Sarvapali D. Ramchurn, Shoaib Ehsan
机构
*
School of Electronics and Computer Science, University of Southampton(南安普顿大学电子与计算机科学学院)
;
School of Electrical Engineering and Computer Science, Queensland University of Technology(昆士兰科技大学电气工程与计算机科学学院)
;
School of Computer Science and Electronic Engineering, University of Essex(埃塞克斯大学计算机科学与电子工程学院)
RoboDesign1M: A Large-scale Dataset for Robot Design Understanding
RoboDesign1M:用于机器人设计理解的大规模数据集
Tri Le, Toan Nguyen, Quang Tran, Quang Nguyen, Baoru Huang, Hoan Nguyen, Minh Nhat Vu, Tung D. Ta, Anh Nguyen
机构
*
FPT Software AI Center(FPT软件人工智能中心)
;
University of Liverpool(利物浦大学)
;
University of Information Technology(信息技术大学)
;
Automation & Control Institute(自动化与控制研究所)
;
Department of Creative Informatics(创意信息系)
;
Faculty of Environment and Information Studies(环境与信息科学系)
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations
XR-1:通过学习统一的视觉-运动表示实现多功能的视觉-语言-动作模型
Shichao Fan, Kun Wu, Zhengping Che, Xinhua Wang, Di Wu, Fei Liao, Ning Liu, Yixue Zhang, Zhen Zhao, Zhiyuan Xu, Meng Li, Qingjie Liu, Shanghang Zhang, Min Wan, Jian Tang
机构
*
Beijing Innovation Center of Humanoid Robotics, Beijing, China(北京人形机器人创新中心,北京,中国)
;
School of Mechanical Engineering and Automation, Beihang University, Beijing, China(北京航空航天大学机械工程及自动化学院,北京,中国)
;
State Key Laboratory of Virtual Reality Technology and Systems, SCSE, Beihang University, Beijing, China(虚拟现实技术与系统国家重点实验室,SCSE,北京航空航天大学,北京,中国)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University, Beijing, China(多媒体信息处理国家重点实验室,计算机科学学院,北京大学,北京,中国)
SilvaScenes: Tree Detection and Species Classification from Under-Canopy Images in Natural Forests
SilvaScenes:自然森林冠层下图像中的树木检测与物种分类
David-Alexandre Duclos, William Guimont-Martin, Gabriel Jeanson, Arthur Larochelle-Tremblay, Martine Lapointe, Théo Defosse, Frédéric Moore, Philippe Nolet, François Pomerleau, Philippe Giguère
机构
*
Northern Robotics Laboratory, Université Laval(北方机器人实验室,拉瓦尔大学)
;
Département des sciences du bois et de la forêt, Université Laval(林业与木材科学系,拉瓦尔大学)
;
Institut des Sciences de la Forêt tempérée, Université du Québec en Outaouais(温带森林科学研究所,魁北克outsouais大学)
NeuralActuator: Neural Actuation Modeling for Robot Dynamics and External Force Perception
神经执行器:用于机器人动力学和外力感知的神经驱动建模
Zhiyang Dou, John U. Onyemelukwe, Hangxing Zhang, Heng Zhang, Minghao Guo, Yunsheng Tian, Michal Piotr Lipiec, Joshua Jacob, Chao Liu, Peter Yichen Chen, Yuri Ivanov, Wojciech Matusik