VisioPath: Vision-Language Enhanced Model Predictive Control for Safe Autonomous Navigation in Mixed Traffic
Shanting Wang, Panagiotis Typaldos, Chenjun Li, Andreas A. Malikopoulos
机构
*
System Engineering Program, Cornell University(Cornell大学系统工程项目)
;
School of Civil and Environmental Engineering, Cornell University(Cornell大学土木与环境工程学院)
;
School of Electrical and Computer Engineering, Cornell University(Cornell大学电气与计算机工程学院)
DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory
Zihe Ji, Huangxuan Lin, Yue Gao
机构
*
SJTU Paris Elite Institute of Technology, Shanghai Jiao Tong University(上海交通大学巴黎精英技术研究所)
;
Department of Automation, Shanghai Jiao Tong University(上海交通大学自动化系)
;
MoE Key Lab of Artificial Intelligence and AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能与AI研究所)
Comments6 figures, 2 tables, Accepted to Robotics: Science and Systems (RSS) 2025 Workshop on Robot Planning in the Era of Foundation Models (FM4RoboPlan)
Commentsaccepted to International Journal of Robotics Research (IJRR). 24 pages, 18 figures. The paper contains texts from VLMaps(arXiv:2210.05714) and AVLMaps(arXiv:2303.07522). The project page is https://mslmaps.github.io/
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
Samsung R&D Institute China–Beijing(三星中国北京研发中心)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
机构
*
Intelligent Space Robotics Laboratory, Center for Digital Engineering, Skolkovo Institute of Science and Technology(智能空间机器人实验室,数字工程中心,斯克尔科沃科学与技术研究所)
专题命中
GUI与屏幕智能体
:VLM(abstract);visual language model(abstract)
CommentsarXiv admin note: text overlap with arXiv:2501.05014
Comments8 pages, 5 figures. This paper has been accepted for publication at the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2024