US-VLA: An Ultrasound Vision-Language-Action Model for Embodied Abdomina
US-VLA:用于腹部超声的超声-视觉-动作模型
机构 * Faculty of Computer Science and Technology, Ocean University of China(中国海洋大学计算机科学与技术学院) ; School of Information Science and Engineering, Shandong University(山东大学信息科学与工程学院) ; the Qilu Second Hospital of Shandong University(山东大学第二齐鲁医院) ; Innovation School of Artificial Intelligence, Hefei University of Technology(合肥工业大学人工智能创新学院)
专题命中 VLA模型 :VLA(title,title_cn);vision-language-action(title,abstract);action model(title,abstract);分类 cs.RO、cs.CV
AI总结 针对现有超声扫描方法泛化性与稳定性不足的问题,提出US-VLA模型,设计超声感知专家融合模块并构建US-VLA-Data数据集,在腹部超声探头操作任务中取得竞争力性能。