WaterVideoQA: ASV-Centric Perception and Rule-Compliant Reasoning via Multi-Modal Agents
WaterVideoQA: 以ASV为中心的感知与符合规则的推理 via 多模态智能体
机构 * Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou)(香港理工大学(广州)人工智能研究所) ; Hubei Key Laboratory of Inland Shipping Technology (Wuhan University of Technology)(湖北内河航运技术重点实验室(武汉理工大学)) ; School of Navigation, Wuhan University of Technology(武汉理工大学航海学院) ; School of Advanced Technology, Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学先进科技学院) ; School of Artificial Intelligence, Nanjing University(南京大学人工智能学院) ; School of Information Engineering, Yancheng Institute of Technology(盐城职业技术学院信息工程学院) ; School of Engineering, Stanford University(斯坦福大学工程学院) ; Centre for AI and Data Science Innovation and the School of Science and Engineering, James Cook University(詹姆斯库克大学人工智能与数据科学创新中心及科学与工程学院)
专题命中 逻辑推理 :reasoning(title,abstract)
AI总结 WaterVideoQA通过多模态智能体系统,实现ASV在复杂水域环境中的感知与规则合规推理,提升自主航行的安全性和精确性。
Comments 11 pages,8 figures