DGSeg: Dynamic Gating of Semantic-Spatial Guided Predictions for Reasoning Segmentation
DGSeg:用于推理分割的语义-空间引导预测的动态门控
Ruizhe Zeng, Siyu Cao, Lu Zhang, Zhiyong Liu
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所多模态人工智能系统国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Nanjing Artificial Intelligence Research of IA(中国科学院自动化所南京人工智能研究院)
Are Neuro-Inspired Multi-Modal Vision-Language Models Resilient to Membership Inference Privacy Leakage?
受神经启发的多模态视觉-语言模型对成员推断隐私泄露是否具有弹性?
David Amebley, Sayanton Dibbo
机构
*
The University of Alabama(阿拉巴马大学)
;
Alabama Center for the Advancement of AI(阿拉巴马人工智能 advancement 中心)
;
Trustworthy AI Lab(可信人工智能实验室)
;
Department of Computer Science, The University of Alabama(计算机科学系)
DenseMLLM: Standard Multimodal LLMs for Dense Prediction
DenseMLLM:用于密集预测的标准多模态大语言模型
Yi Li, Hongze Shen, Lexiang Tang, Xin Li, Xinpeng Ding, Yinsong Liu, Deqiang Jiang, Xing Sun, Xiaomeng Li
机构
*
Department of Electronic and Computer Engineering, The Hong Kong University of Science and Technology, Hong Kong, China(香港科技大学电子与计算机工程系)
;
Tencent, Youtu-Lab, China(腾讯优图实验室)