Geometrical Cross-Attention and Nonvoid Voxelization for Efficient 3D Medical Image Segmentation
几何交叉注意力与非空体素化用于高效3D医学图像分割
Chenxin Yuan, Shoupeng Chen, Haojiang Ye, Yiming Miao, Limei Peng, Pin-Han Ho
机构
*
organization= Shenzhen Institute for Advanced Study , addressline= University of Electronic Science
;
organization= School of Electrical
;
Electronic Engineering , addressline= Nanyang Technological University , city= Singapore , postcode= 639798 , country= Singapore
;
organization= School of Science
;
Engineering , addressline= The Chinese University of Hong Kong, Shenzhen , city= Shenzhen , postcode= 518172 , state= Guangdong , country= China
;
organization= School of Computer Science
;
Engineering , addressline= Kyungpook National University , city= Daegu , postcode= 37224 , country= South Korea
;
organization= Department of Electrical
;
Computer Engineering , addressline= University of Waterloo , city= Waterloo , postcode= N2L3G1 , state= Ontario , country= Canada
Integration of Object Detection and Small VLMs for Construction Safety Hazard Identification
对象检测与小型视觉语言模型的整合用于建筑安全危险识别
Muhammad Adil, Mehmood Ahmed, Muhammad Aqib, Vicente A. Gonzalez, Gaang Lee, Qipei Mei
机构
*
Infrastructure Human Tech (IHT) Lab, Department of Civil and Environmental Engineering, University of Alberta, Edmonton, Alberta, Canada(基础设施人类技术(IHT)实验室,土木与环境工程系,阿尔伯塔大学,埃德蒙顿,阿尔伯塔,加拿大)
Beyond the Global Scores: Fine-Grained Token Grounding as a Robust Detector of LVLM Hallucinations
超越全局评分:细粒度令牌接地作为检测LVLM幻觉的稳健检测器
Tuan Dung Nguyen, Minh Khoi Ho, Qi Chen, Yutong Xie, Nguyen Cam-Tu, Minh Khoi Nguyen, Dang Huy Pham Nguyen, Anton van den Hengel, Johan W. Verjans, Phi Le Nguyen, Vu Minh Hieu Phan
机构
*
Hanoi University of Science and Technology(河内科技大学)
;
Australian Institute for Machine Learning, University of Adelaide(澳大利亚机器学习研究所,阿德莱德大学)
;
Nanjing University(南京大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Hanoi-Amsterdam High School for the Gifted(河内阿姆斯特丹天才高中)
Benchmarking and Evaluating VLMs for Software Architecture Diagram Understanding
对软件架构图理解的VLMs基准测试与评估
Shuyin Ouyang, Jie M. Zhang, Jingzhi Gong, Gunel Jahangirova, Mohammad Reza Mousavi, Jack Johns, Beum Seuk Lee, Adam Ziolkowski, Botond Virginas, Joost Noppen
Journal refProceedings of the 21st International Conference on Computer Vision Theory and Applications - Volume 3: VISAPP 2026; ISBN 978-989-758-804-4; ISSN 2184-4321, SciTePress, pages 353-364
ArchMap: Arch-Flattening and Knowledge-Guided Vision Language Model for Tooth Counting and Structured Dental Understanding
ArchMap:用于牙齿计数和结构牙科理解的拱形扁平化和知识引导的视觉语言模型
Bohan Zhang, Yiyi Miao, Taoyu Wu, Tong Chen, Ji Jiang, Zhuoxiao Li, Zhe Tang, Limin Yu, Jionglong Su
机构
*
Xi'an Jiaotong-Liverpool University(西交利物浦大学)
;
University of Liverpool(利物浦大学)
;
Zhejiang University of Technology(浙江工业大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
机构
*
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院)
;
College of Information Science and Technology, Eastern Institute of Technology(东方理工学院信息科学与技术学院)