Task-Adaptive 3D Cross-Field MRI Translation via Field-Conditioned Content-Style Pretraining
基于场条件内容-风格预训练的任务自适应3D跨场MRI转换
Haowen Pang, Yingqi Hao, Pengli Zhu
机构
*
School of Integrated Circuits and Electronics, Beijing Institute of Technology(北京理工大学集成电路与电子学院)
;
School of Biomedical Engineering, Tsinghua Medicine, Tsinghua University(清华大学医学院生物医学工程学院)
;
Department of Electronic Engineering, The Chinese University of Hong Kong(香港中文大学电子工程系)
CommentsAccepted at the Computer Vision in Plant Phenotyping and Agriculture (CVPPA) Workshop at the European Conference on Computer Vision (ECCV) 2026
机构
*
State Key Laboratory of General Artificial Intelligence, Peking University, Shenzhen Graduate School(国家一般人工智能重点实验室,北京大学深圳研究生院)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究所,中国科学院)
;
State Key Laboratory of Internet of Things for Smart City, University of Macau(智慧城市物联网重点实验室,澳门大学)
Quality Text, Robust Vision: The Role of Language in Enhancing Visual Robustness of Vision-Language Models
高质量文本,稳健视觉:语言在增强视觉语言模型视觉稳健性中的作用
Futa Waseda, Saku Sugawara, Isao Echizen
机构
*
The University of Tokyo(东京大学)
;
National Institute of Informatics(日本信息处理研究所)
;
The University of Tokyo, National Institute of Informatics(东京大学、日本信息处理研究所)
TokenSwap: Backdoor Attack on the Compositional Understanding of Large Vision-Language Models
TokenSwap:对大型视觉语言模型组合理解的后门攻击
Zhifang Zhang, Qiqi Tao, Jiaqi Lv, Na Zhao, Lei Feng, Joey Tianyi Zhou
机构
*
School of Computer Science and Engineering, Southeast University, China(东南大学计算机科学与工程学院)
;
School of Electrical Engineering and Computer Science, University of Queensland, Australia(昆士兰大学电气工程与计算机科学学院)
;
Design Pillar, Singapore University of Technology and Design(新加坡科技设计大学设计学院)
;
Centre for Frontier AI Research (CFAR), Agency for Science, Technology and Research (A*STAR), Singapore(前沿人工智能研究中心(CFAR))
;
Institute of High Performance Computing (IHPC), Agency for Science, Technology and Research (A*STAR), Singapore(高性能计算研究所(IHPC))
Back to Point: Exploring Point-Language Models for Zero-Shot 3D Anomaly Detection
回到点:探索用于零样本3D异常检测的点语言模型
Kaiqiang Li, Gang Li, Mingle Zhou, Min Li, Delong Han, Jin Wan
机构
*
Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center (National Supercomputer Center in Jinan), Qilu University of Technology (Shandong Academy of Sciences), Jinan, China(计算机功率网络与信息安全重点实验室,教育部,山东计算机科学中心(济南国家超级计算机中心),齐鲁大学(山东科学院),济南,中国)
;
Shandong Provincial Key Laboratory of Computing Power Internet and Service Computing, Shandong Fundamental Research Center for Computer Science, Jinan, China(山东省计算功率互联网与服务计算重点实验室,山东省计算机科学基础研究中心,济南,中国)
LA4VLA: Learning to Act without Seeing via Language-Action Pretraining
LA4VLA:通过语言-动作预训练实现无视觉行动学习
Tao Lin, Yuxin Du, Yiran Mao, Zewei Ye, Yilei Zhong, Bing Cheng, Yiming Wang, Jiting Liu, Yang Tian, Junchi Yan, Feiran Wu, Zenan Meng, Hu Wei, Yuqian Fu, Gen Li, Bo Zhao
机构
*
School of AI, Shanghai Jiao Tong University(上海交通大学人工智能学院)
;
Alibaba Group(阿里巴巴集团)
;
Nanyang Technological University(南洋理工大学)
;
KAUST(阿卜杜拉国王科技大学)
Technical Report for the ICRA 2026 GOOSE 2D Fine-Grained Semantic Segmentation Challenge: Pretraining-Diverse Ensemble of Foundation Vision Encoders for Robust Outdoor Scene Understanding
ICRA 2026 GOOSE 2D细粒度语义分割挑战赛技术报告:面向鲁棒户外场景理解的预训练多样化基础视觉编码器集成