GroundingME: Exposing the Visual Grounding Gap in MLLMs through Multi-Dimensional Evaluation
GroundingME:通过多维评估揭示MLLMs中的视觉 grounding 隙
Rang Li, Lei Li, Shuhuai Ren, Hao Tian, Shuhao Gu, Shicheng Li, Zihao Yue, Yudong Wang, Wenhan Ma, Zhe Yang, Jingyuan Ma, Zhifang Sui, Fuli Luo
机构
*
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机学院,北京大学)
;
LLM-Core Xiaomi(小米LLM-Core)
;
The University of Hong Kong(香港大学)
;
Renmin University of China(中国人民大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract)
GIR-Bench: Versatile Benchmark for Generating Images with Reasoning
GIR-Bench:用于生成图像的多功能基准
Hongxiang Li, Yaowei Li, Bin Lin, Yuwei Niu, Yuhang Yang, Xiaoshuang Huang, Jiayin Cai, Xiaolong Jiang, Yao Hu, Long Chen
机构
*
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Peking University(北京大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Xiaohongshu Inc.(小红书公司)
专题命中
评测与基准
:large language model(abstract);language model(abstract)
M4-RAG: A Massive-Scale Multilingual Multi-Cultural Multimodal RAG
M4-RAG:大规模多语言多文化多模态检索增强生成
David Anugraha, Patrick Amadeus Irawan, Anshul Singh, En-Shiun Annie Lee, Genta Indra Winata
机构
*
Stanford University(斯坦福大学)
;
MBZUAI
;
Indian Institute of Science(印度科学研究院)
;
Ontario Tech University(安大略技术大学)
;
University of Toronto(多伦多大学)
;
Capital One
Multiperspectivity as a Resource for Narrative Similarity Prediction
多视角作为叙事相似性预测的资源
Max Upravitelev, Veronika Solopova, Jing Yang, Charlott Jakob, Premtim Sahitaj, Ariana Sahitaj, Vera Schmitt
机构
*
Technische Universit‘̀at Berlin(柏林技术大学)
;
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)
;
BIFOLD – Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究院)
;
Centre for European Research in Trusted AI (CERTAIN)(可信人工智能欧洲研究中心)
Entropy Alone is Insufficient for Safe Selective Prediction in LLMs
仅凭熵不足以在LLMs中实现安全的选择性预测
Edward Phillips, Fredrik K. Gustafsson, Sean Wu, Anshul Thakur, David A. Clifton
机构
*
Department of Engineering Science, University of Oxford(牛津大学工程科学系)
;
Oxford Suzhou Centre for Advanced Research, University of Oxford(牛津大学苏州市先进研究中心)
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学)
;
Faculty of Computility Microelectronics, Shenzhen University of Advanced Technology and the Guangdong Provincial Key Laboratory of Computility Microelectronics(计算微电子学系,深圳先进技术大学及广东省计算微电子重点实验室)
;
College of Software Engineering, Xi’an Jiaotong University(软件工程学院,西安交通大学)