A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy
胃肠内窥镜中视觉语言模型幻觉检测的基准测试
Aminu Lawal, Niyoj Oli, Sachin Acharya, Prashnna Gyawali, Maria Carmen Romano, Binod Bhattarai
机构
*
University of Aberdeen(阿伯丁大学)
;
Nepal Applied Mathematics and Informatics Institute for Research(尼泊尔应用数学与信息学研究所)
;
West Virginia University(西弗吉尼亚大学)
HANCLIP: A Family of Hyperbolic Angular Negation Vision Language Models
HANCLIP:双曲角否定视觉语言模型系列
Hoang-Bao Le, Aiden Durrant, Thai Son Mai, Binh T. Nguyen, Liting Zhou, Cathal Gurrin
机构
*
ADAPT Centre Dublin City University, Ireland(爱尔兰都柏林城市大学ADAPT中心)
;
University of East Anglia Norwich, UK(英国东英吉利大学)
;
Queen’s University Belfast Belfast, UK(英国贝尔法斯特女王大学)
;
University of Science Vietnam National University Ho Chi Minh City, Vietnam(越南胡志明市国家大学理科大学)
专题命中
视觉推理
:vision language model(title);vision-language model(abstract);VLM(abstract_cn);分类 cs.CV
Neuro-Symbolic Drive: Rule-Grounded Faithful Reasoning for Driving VLAs
神经符号驱动:基于规则忠实推理的驾驶VLA
Xiangbo Gao, Xiukun Huang, Boyu Lu, Junge Zhang, Mengjie Mao, Jiachen Li, Wei Xiong, Zhengzhong Tu
机构
*
Texas A&M University(德克萨斯农工大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Maryland(马里兰大学)
;
University of California, Riverside(加利福尼亚大学河滨分校)
;
University of Pittsburgh(匹兹堡大学)
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴达摩院)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Monash University(莫纳什大学)
;
TRE, Alibaba Group(阿里巴巴TRE)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Xiamen University(厦门大学)
;
Kling Team, Kuaishou Technology(快手科技Kling团队)
;
National University of Singapore(新加坡国立大学)
;
Southern University of Science and Technology(南方科技大学)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.AI
UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving
UniDrive: 面向自动驾驶可解释风险理解的统一视觉-语言与定位框架
Xiaowei Gao, Pengxiang Li, Yitai Cheng, Ruihan Xu, James Haworth, Stephen Law, Yun Ye
机构
*
organization= Department of Earth Science \& Engineering, Imperial College London , city= London , postcode= SW7 2AZ , country= United Kingdom
;
organization= SpaceTimeLab, Department of Civil, Environmental
;
Geomatic Engineering, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Department of Computing, The Hong Kong Polytechnic University , city= Hong Kong , country= China
;
organization= Trinity College, University of Oxford , city= Oxford , postcode= OX1 3BH , country= United Kingdom
;
organization= Department of Geography, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom
;
organization= Centre for Global Infrastructure Resilience, The Bartlett School of Sustainable Construction, University College London , city= London , postcode= WC1E 7HB , country= United Kingdom
专题命中
视觉定位与Grounding
:grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI
CommentsAccepted at FOIS 2026 (16th International Conference on Formal Ontology in Information Systems), Vitória, Brazil; to appear in Frontiers in Artificial Intelligence and Applications, IOS Press. 16 pages, 1 figure, 2 tables
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴集团达摩院)
;
Hupan Lab(华平实验室)
;
Huazhong University of Science and Technology(华中科技大学)
;
East China Normal University(华东师范大学)
;
Shanghai Jiao Tong University(上海交通大学)
机构
*
University of New South Wales(新南威尔士大学)
;
National University of Singapore(新加坡国立大学)
;
NVIDIA(英伟达)
;
Nanjing University(南京大学)
;
University of Technology Sydney(悉尼科技大学)
;
Australian National University(澳大利亚国立大学)
机构
*
Institute of Foundation Models, Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学基础模型研究所)
;
School of Computer Science, Carnegie Mellon University(卡内基梅隆大学计算机科学学院)
机构
*
Tsinghua University(清华大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Tsinghua Shenzhen International Graduate School(清华大学深圳国际研究生院)
;
Dalian University of Technology(大连理工大学)