Improving Large Vision-Language Models' Understanding for Flow Field Data
提升大型视觉-语言模型对流场数据的理解能力
机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
专题命中 视觉问答 :vision-language model(title,abstract);visual question answering(abstract);分类 cs.CV
AI总结 FieldLVLM通过场感知语言生成策略和数据压缩多模态模型调优,提升大型视觉-语言模型对流场数据的理解能力。
Comments Accepted by Machine Intelligence Research