Search-based Testing of Vision Language Models for In-Car Scene Understanding
基于搜索的车内场景理解视觉语言模型测试
Lev Sorokin, Chen Yang, Ken E. Friedl, Andrea Stocco
机构
*
BMW Group, Technical University of Munich(宝马集团、慕尼黑技术大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Technical University of Munich, fortiss GmbH(慕尼黑技术大学、fortiss GmbH)
WING: A Window-Prior-Based Generative Network with Gated Inception for Cross-Modality CT Synthesis
WING:一种基于窗口先验的带门控inception的生成网络用于跨模态CT合成
Siyuan Mei, Yan Xia, Yipeng Sun, Siming Bayer, Zirong Li, Chengze Ye, Daiqi Liu, Fuxin Fan, Yixing Huang, Andreas Maier
机构
*
Pattern Recognition Lab, Friedrich-Alexander-Universität Erlangen-Nürnberg(模式识别实验室,埃尔朗根-纽伦堡弗里德里希-亚历山大大学)
;
Department of Orthodontics and Orofacial Orthopaedics, Friedrich-Alexander-Universität Erlangen-Nürnberg(正畸与口腔颌面正畸科,埃尔朗根-纽伦堡弗里德里希-亚历山大大学)
;
Siemens Healthineers(西门子医疗)
;
Institute of Medical Technology, Peking University(北京大学医学技术研究所)
机构
*
Department of Gynaecology and Obstetrics, The Affiliated Jiangyin Hospital of Nantong University(南通大学附属江阴医院妇产科)
;
Department of Oncology, the Affiliated Jiangyin Hospital of Nantong University(南通大学附属江阴医院肿瘤科)
;
Thrust of AI, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)人工智能推力实验室)
;
FertiTech AI(生殖科技人工智能公司)
;
Department of Oncology, Suzhou Xiangcheng People’s Hospital(苏州市相城区人民医院肿瘤科)
;
Department of Biological Sciences and Bioinformatics, School of Science, Xi’an Jiaotong-Liverpool University(西交利物浦大学理学院生物科学与生物信息学系)
;
School of Artificial Intelligence and Computer Science, Nantong University(南通大学人工智能与计算机科学学院)
MergeSurv: Merging-Based Continual Learning for Survival Analysis on Whole-Slide Images
MergeSurv:用于全切片图像生存分析的基于合并的持续学习
Vu Minh Tran, Doanh C. Bui, Maï K. Nguyen, Khang Nguyen
机构
*
University of Information Technology(信息技术大学)
;
Viet Nam National University Ho Chi Minh City(胡志明市越南国立大学)
;
ETIS (UMR 8051), CY Cergy Paris University, ENSEA, CNRS, France(法国CY Cergy巴黎大学、ENSEA、法国国家科学研究中心联合建立的ETIS实验室(UMR 8051))
机构
*
Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)
;
Department of Thoracic Surgery, Peking University People’s Hospital(北京大学人民医院胸外科)
;
Thoracic Oncology Institute, Peking University People’s Hospital(北京大学人民医院胸部肿瘤研究所)
;
Research Unit of Intelligence Diagnosis and Treatment in Early Non-small Cell Lung Cancer, Chinese Academy of Medical Sciences(中国医学科学院早期非小细胞肺癌智能诊断与治疗研究组)
;
Institute of Advanced Clinical Medicine, Peking University(北京大学先进临床医学院)
;
Beijing Key Laboratory of Innovative Application of Big Data in Lung Cancer, Peking University People’s Hospital(北京大学人民医院肺癌大数据创新应用北京市重点实验室)
Learning to Watch: Active Video Anomaly Understanding via Interleaved Policy Optimization
学习观察:通过交错策略优化的主动视频异常理解
Mengjingcheng Mo, Jiaxu Leng, Xinbo Gao
机构
*
School of Computer Science and Technology, Chongqing University of Posts and Telecommunications, Chongqing, China(重庆邮电大学计算机科学与技术学院)
;
School of Computer Science(计算机科学学院)
;
Chongqing College of Artificial Intelligence, Chongqing, China(重庆人工智能学院)
CommentsThis work is accepted by CVPR'26, Embodied AI Workshop. This paper represent a part of early result of our official world-action model zero-shot sim-to-real transfer work, which will be released soon
Multi-Modal Conditioned High-Resolution Transformer for Urban Electromagnetic Field Map Prediction Download PDF
面向城市电磁场地图预测的多模态条件高分辨率Transformer
Do-Eon Kim, Dongryul Park, Seungyoung Ahn, Namwoo Kang, Seong-heum Kim, Seongsin Kim
机构
*
Soongsil University(崇实大学)
;
Cho Chun Shik Graduate School of Mobility, Korea Advanced Institute of Science and Technology(韩国科学技术院赵春植移动研究生院)
;
Department of Intelligent Semiconductors, Soongsil University(崇实大学智能半导体系)