MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification
用于鲁棒跨数据集图像分类的MLLM路由异质集成模型
机构 * The Bredesen Center for Interdisciplinary Research and Graduate Education(布雷德森跨学科研究与研究生教育中心) ; University of Tennessee Knoxville(田纳西大学诺克斯维尔分校)
专题命中 视觉推理 :MLLM(title,title_cn);vision-language model(abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI、cs.LG
AI总结 该研究针对跨数据集图像分类泛化难题,提出ARMDIL模型,通过MLLM智能体动态路由图像到适配的视觉骨干,兼具竞争力、高适应性与可解释性,为通用视觉系统奠定基础。
Comments 8 pages, 4 figures, 7 tables