机构
*
Technical University of Munich(慕尼黑技术大学)
;
Helmholtz Munich(亥姆霍兹慕尼黑)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
Munich Data Science Institute(慕尼黑数据科学研究所)
;
University of Tübingen(图宾根大学)
;
University of Copenhagen(哥本哈根大学)
机构
*
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
;
Zhongtai Securities Institute for Financial Studies, Shandong University(山东大学中泰证券金融研究学院)
DepthVision: Enabling Robust Vision-Language Models with GAN-Based LiDAR-to-RGB Synthesis for Autonomous Driving
Sven Kirchner, Nils Purschke, Ross Greer, Alois C. Knoll
机构
*
Chair of Robotics, Artificial Intelligence and Real-time Systems, Technical University of Munich(机器人学、人工智能与实时系统教授会,慕尼黑技术大学)
;
Computer Science and Engineering Department, University of California Merced(计算机科学与工程系,加州大学默塞德分校)
HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
Sushant Gautam, Michael A. Riegler, Pål Halvorsen
机构
*
Simula Metropolitan Center for Digital Engineering (SimulaMet)(Simula数字工程研究中心)
;
Oslo Metropolitan University (OsloMet)(奥斯陆 Metropolitan 大学)
;
Simula Research Laboratory(Simula研究实验室)
机构
*
Hong Kong University of Science and Technology(香港科技大学)
;
Kyoto University(京都大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
The University of Texas at Dallas(德克萨斯大学达拉斯分校)
;
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
Cornell University(康奈尔大学)
;
Indiana University(印第安纳大学)
;
National Taiwan Normal University(台湾师范大学)
;
University of Liverpool(利物浦大学)
;
University of Edinburgh(爱丁堡大学)
;
Zhejiang University(浙江大学)
;
Purdue University(普渡大学)
;
Emory University(埃默里大学)
;
West China Biomedical Big Data Center, West China Hospital, Sichuan University(西京生物大数据中心,西京医院,四川大学)
Generating Accurate and Detailed Captions for High-Resolution Images
Hankyeol Lee, Gawon Seo, Kyounggyu Lee, Dogun Kim, Kyungwoo Song, Jiyoung Jung
机构
*
Department of Artificial Intelligence, University of Seoul(首尔大学人工智能系)
;
Department of Computer Science and Engineering, POSTECH(POSTECH计算机科学与工程系)
;
Department of Applied Statistics, Yonsei University(延世大学应用统计系)
专题命中
图文多模态
:multimodal(abstract);分类 cs.CV、cs.AI
CommentsWork conducted in 2024; released for archival purposes