AeroScene: Progressive Scene Synthesis for Aerial Robotics
AeroScene:用于空域机器人领域的渐进式场景合成
Nghia Vu, Tuong Do, Dzung Tran, Binh X. Nguyen, Hoan Nguyen, Erman Tjiputra, Quang D. Tran, Hai-Nguyen Nguyen, Anh Nguyen
机构
*
University of Liverpool, UK(英国利物浦大学)
;
AIOZ Ltd., Singapore(新加坡AIOZ有限公司)
;
National Tsing Hua University, Taiwan(台湾国立清华大学)
;
RMIT University, Vietnam Campus(越南RMIT大学校园)
;
University of Information Technology, VNUHCM, Vietnam(越南VNUHCM大学信息科技学院)
Comments23 pages, Keywords: Language Grounding, Language Granularity, Instruction Following Agent, Width-based Planning Research Area: Multimodality and Language Grounding to Vision, Robotics and Beyond Research Area Keywords: vision language navigation, multimodality, neurosymbolic approaches
Comments26 pages, 13 Figures, 4 Tables. Revised manuscript with a clearer state-of-the-art discussion, reorganized methodology, and updated figures and content
EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild
EgoWalk:一种用于野外机器人导航的多模态数据集
Timur Akhtyamov, Mohamad Al Mdfaa, Javier Antonio Ramirez Benavides, Arthur Nigmatzyanov, Sergey Bakulin, German Devchich, Denis Fatykhov, Diego Ruiz Salinas, Alexander Mazurov, Kristina Zipa, Malik Mohrat, Pavel Kolesnik, Ivan Sosin, Gonzalo Ferrer
机构
*
Skolkovo Institute of Science and Technology(斯克尔科沃科学与技术研究所)
;
Sber Robotics Center(Sber机器人中心)
Model-free source seeking of exponentially convergent unicycle: theoretical and robotic experimental results
无模型的指数收敛单轮车源寻找:理论和机器人实验结果
Rohan Palanikumar, Ahmed A. Elgohary, Victoria Grushkovskaya, Sameh A. Eisa
机构
*
Department of Aerospace Engineering and Engineering Mechanics(航空航天工程与工程力学系)
;
University of Cincinnati(辛辛那提大学)
;
Department of Mathematics(数学系)
;
University of Klagenfurt(克雷格弗特大学)
COFFAIL: A Dataset of Successful and Anomalous Robot Skill Executions in the Context of Coffee Preparation
COFFAIL:咖啡准备情境中成功与异常机器人技能执行数据集
Alex Mitrevski, Ayush Salunke
机构
*
Division of Systems and Control, Chalmers University of Technology(系统与控制系,楚姆勒斯理工大学)
;
Bonn-Rhein-Sieg University of Applied Sciences(波恩-莱茵-西弗大学)
;
Autonomous Systems Group, Bonn-Rhein-Sieg University of Applied Sciences(自主系统组,波恩-莱茵-西弗大学)
Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction
交互链基准(COIN):当推理遇见具身交互
Xianhao Wang, Xiaojian Ma, Haozhe Hu, Rongpeng Su, Yutian Cheng, Zhou Ziheng, Hangxin Liu, Lei Liu, Bin Li, Qing Li
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Beijing Institute for General Artificial Intelligence (BIGAI)(北京通用人工智能研究院)
;
Xidian University(西安电子科技大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
A VLM-based Method for Visual Anomaly Detection in Robotic Scientific Laboratories
基于VLM的方法用于机器人科学实验室中的视觉异常检测
Shiwei Lin, Chenxu Wang, Xiaozhen Ding, Yi Wang, Boyuan Du, Lei Song, Chenggang Wang, Huaping Liu
机构
*
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
School of Physics and Electronic Information, Yantai University(烟台大学物理与电子信息学院)
;
School of Computer and Big Data, Fuzhou University(福州大学计算机与大数据学院)
;
Department of Automation, Shanghai Jiao Tong University(上海交通大学自动化系)
Using large language models for embodied planning introduces systematic safety risks
使用大型语言模型进行具身规划引入了系统性的安全风险
Tao Zhang, Kaixian Qu, Zhibin Li, Jiajun Wu, Marco Hutter, Manling Li, Fan Shi
机构
*
ETH Zurich(苏黎世联邦理工学院)
;
University College London(伦敦大学学院)
;
Stanford University(斯坦福大学)
;
Northwestern University(西北大学)
;
National University of Singapore(新加坡国立大学)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
R3D2:通过扩散实现自动驾驶模拟中的真实3D资产插入
William Ljungbergh, Bernardo Taveira, Wenzhao Zheng, Adam Tonderski, Chensheng Peng, Fredrik Kahl, Christoffer Petersson, Michael Felsberg, Kurt Keutzer, Masayoshi Tomizuka, Wei Zhan
Bridging the Ex-Vivo to In-Vivo Gap: Synthetic Priors for Monocular Depth Estimation in Specular Surgical Environments
弥合体外到体内差距:用于镜面外科环境单目深度估计的合成先验
Ankan Aich, Emma D. Ryan, Kris Moe, Isaac Schmale, Li-Xing Man, Yangming Lee
机构
*
RoCAL, Rochester Institute of Technology(罗切斯特理工学院RoCAL实验室)
;
Department of Otolaryngology Head and Neck Surgery, University of Rochester Medical Center(罗切斯特大学医学中心耳鼻喉科与头颈外科部门)
;
University of Washington, Department of Otolaryngology–Head and Neck Surgery(华盛顿大学耳鼻喉科与头颈外科部门)
机构
*
Xi’an Jiaotong University(西安交通大学)
;
Hefei University of Technology(合肥工业大学)
;
CSIRO(澳大利亚联邦科学与工业研究组织)
;
Northwestern Polytechnical University(西北工业大学)
;
University of Macau(澳门大学)
RAYEN: Imposition of Hard Convex Constraints on Neural Networks
RAYEN:在神经网络中施加硬凸约束
Jesus Tordesillas, Victor Klemm, Jonathan P. How, Marco Hutter
机构
*
Institute for Research in Technology, ICAI School of Engineering, Comillas Pontifical University(技术研究 institute,ICAI 工程学院,Comillas 大学)
;
Aerospace Controls Laboratory, Massachusetts Institute of Technology(航空航天控制实验室,麻省理工学院)
Fringe Projection Based Vision Pipeline for Autonomous Hard Drive Disassembly
基于条纹投影的自主硬盘拆解视觉流水线
Badrinath Balasubramaniam, Vignesh Suresh, Benjamin Metcalf, Beiwen Li
机构
*
School of Electrical and Computer Engineering, University of Georgia(佐治亚大学电气与计算机工程学院)
;
Alcon Research Laboratories(阿康研究实验室)
;
School of Environmental, Civil, Agricultural and Mechanical Engineering, University of Georgia(佐治亚大学环境、土木、农业和机械工程学院)
Camo-M3FD: A New Benchmark Dataset for Cross-Spectral Camouflaged Pedestrian Detection
Camo-M3FD:一种新的跨光谱伪装行人检测基准数据集
Henry O. Velesaca, Andrea Mero, Guillermo A. Castillo, Angel D. Sappa
机构
*
ESPOL Polytechnic University(ESPOL理工大学)
;
Computer Vision Center, Universitat Autònoma de Barcelona(巴塞罗那自治大学计算机视觉中心)
;
Software Engineering Department, Research Center for Information and Communication Technologies (CITIC-UGR), University of Granada(格拉纳达大学软件工程系,信息与通信技术研究中心(CITIC-UGR))
;
Università della Svizzera Italiana(瑞士意大利大学)