VLM-SFD: VLM-Assisted Siamese Flow Diffusion Framework for Dual-Arm Cooperative Manipulation
VLM-SFD:基于视觉语言模型的双臂协作操作Siamese流扩散框架
Jiaming Chen, Yiyu Jiang, Aoshen Huang, Yang Li, Wei Pan
机构
*
Department of Computer Science, The University of Manchester(计算机科学系,曼彻斯特大学)
;
School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学)
UTI-LLM: A Personalized Articulatory-Speech Therapy Assistance System Based on Multimodal Large Language Model
Yudong Yang, Xiaokang Liu, Shaofeng zhao, Rongfeng Su, Nan Yan, Lan Wang
机构
*
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences, China(深圳先进技术研究院,中国科学院,中国)
;
Key Laboratory of Biomedical Imaging Science and System, Chinese Academy of Sciences, China(生物医学成像科学与系统重点实验室,中国科学院,中国)
专题命中
VLM训练与架构
:multimodal large language model(title,abstract);MLLM(abstract)
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
Shakib Yazdani, Yasser Hamidullah, Cristina España-Bonet, Josef van Genabith
机构
*
German Research Center for Artificial Intelligence (DFKI GmbH)(德国人工智能研究中心(DFKI GmbH))
;
Saarland Informatics Campus(萨尔兰州信息技术校区)
;
Barcelona Supercomputing Center (BSC-CNS)(巴塞罗那超级计算中心(BSC-CNS))
专题命中
VLM训练与架构
:vision-language model(title);vision language model(abstract);VLM(abstract)