机构
*
Centennial High School, Frisco, Texas, USA(Centennial High School, Texas, USA)
;
Lebanon Trail High School, Frisco, Texas, USA(Lebanon Trail High School, Texas, USA)
;
West Windsor-Plainsboro High School, Princeton Junction, New Jersey, USA(West Windsor-Plainsboro High School, New Jersey, USA)
;
Algoverse AI Research, Palo Alto, California, USA(Algoververse AI Research, California, USA)
专题命中
视觉推理
:VLM(title_cn);vision-language model(abstract);vision language model(abstract);visual reasoning(abstract)
Comments14 pages, 4 figures, 8 tables. Presented at the 39th Conference on Neural Information Processing Systems Workshop: VLM4RWD. Presented at the 43th International Conference on Machine Learning Workshops: ICML 2026 CTB, ICML 2026 FAGEN, ICML 2026 EMM-QA. Authors Aahana Basappa and Pranay Goel contributed equally. Code: https://github.com/AahanaB24/AMVICC, Data: https://doi.org/10.5281/zenodo.17646068
Laurens Samson, Nimrod Barazani, Sennay Ghebreab, Yuki M. Asano
机构
*
Socially-Intelligent Artificial Systems Group, University of Amsterdam(智能社会人工智能系统组,阿姆斯特丹大学)
;
University of Amsterdam(阿姆斯特丹大学)
;
Fundamental AI Lab, University of Technology Nuremberg(基础人工智能实验室,纽伦堡技术大学)
专题命中
VLM训练与架构
:visual language model(title,abstract);VLM(abstract_cn);分类 cs.CV
MedLayBench-V: A Large-Scale Benchmark for Expert-Lay Semantic Alignment in Medical Vision Language Models
MedLayBench-V:面向医学视觉语言模型中专家与普通人语义对齐的大规模基准
Han Jang, Junhyeok Lee, Heeseong Eum, Kyu Sung Choi
机构
*
Seoul National University(首尔国立大学)
;
Seoul National University College of Medicine(首尔国立大学医学院)
;
Department of Radiology, Seoul National University Hospital(首尔国立大学医院放射科)
;
Healthcare AI Research Institute, Seoul National University Hospital(首尔国立大学医院健康人工智能研究所)
;
The Advanced Imaging and Computational Neuroimaging (AICON) Laboratory(先进影像与计算神经影像实验室)
专题命中
VLM训练与架构
:vision language model(title);vision-language model(abstract)
ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data
ZeroWBC: 从人类自我中心数据学习自然全身人形交互
Haoran Yang, Jiacheng Bao, Yucheng Xin, Haoming Song, Yuyang Tian, Bin Zhao, Dong Wang, Xuelong Li
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Northwestern Polytechnical University(西北工业大学)
;
Tsinghua University(清华大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
TeleAI, China Telecom(TeleAI,中国电信)
GeoRanker: Distance-Aware Ranking for Worldwide Image Geolocalization
GeoRanker:面向全球图像地理定位的距离感知排序
Pengyue Jia, Seongheon Park, Song Gao, Xiangyu Zhao, Sharon Li
机构
*
Department of Data Science, City University of Hong Kong(城市大学数据科学系)
;
Department of Computer Sciences, University of Wisconsin-Madison(威斯康星大学麦迪逊分校计算机科学系)
;
Department of Geography, University of Wisconsin-Madison(威斯康星大学麦迪逊分校地理系)