LAION-5B: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, Patrick Schramowski, Srivatsa Kundurthy, Katherine Crowson, Ludwig Schmidt, Robert Kaczmarczyk, Jenia Jitsev
机构
*
Alibaba Group(阿里巴巴集团)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
Nanjing University(南京大学)
;
Peking University(北京大学)
;
Tsinghua University(清华大学)
机构
*
School of Information Science and Technology, ShanghaiTech University(信息科学与技术学院,上海科技大学)
;
Transcengram
;
DeepSeek AI
;
University of Hong Kong(香港大学)
FlowInOne:Unifying Multimodal Generation as Image-in, Image-out Flow Matching
FlowInOne:将多模态生成统一为图像输入、图像输出的流匹配
Junchao Yi, Rui Zhao, Jiahao Tang, Weixian Lei, Linjie Li, Qisheng Su, Zhengyuan Yang, Lijuan Wang, Xiaofeng Zhu, Alex Jinpeng Wang
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Central South University(中南大学)
;
National University of Singapore(新加坡国立大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Microsoft(微软)
Enhancing Explainable Cardiac Diagnosis with Guide-Grounded Multimodal LLMs
用指南引导的多模态语言模型增强可解释的心脏诊断
Hai-Nam Duy Vuong, Duy-Anh Bui, Trong-Nghia Nguyen, Kim-Ngan Thi Nguyen, Trang Mai Xuan, Tien-Cuong Nguyen, Van-Dem Pham, Thien Van Luong
机构
*
Business AI Lab, College of Technology, National Economics University(商业人工智能实验室,技术学院,越南国家经济大学)
;
A2I Lab, Phenikaa School of Computing, Phenikaa University(A2I实验室,费尼卡计算机学院,费尼卡大学)
;
VNPT AI, VNPT Group(VNPT人工智能,VNPT集团)
;
FPT University(FPT大学)
;
Department of Pediatrics, Hospital of University Medicine and Pharmacy, Vietnam National University Hanoi(越南河内国家大学医学与药学院儿科学系)
机构
*
College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院)
;
Shanghai Research Institute for Intelligent Autonomous Systems, Tongji University(同济大学上海自主智能无人系统科学中心)
MM-TRELLIS: Point-Cloud Guided Multi-Modal 3D Vehicle Generation in Autonomous Driving
MM-TRELLIS: 自动驾驶中基于点云引导的多模态3D车辆生成
Hongli Xiao, Youjian Zhang, Yucai Bai, Chaoyue Wang, Yaohui Jin, Xiaoguang Ren, Wenjing Yang, Long Lan
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院教育部人工智能重点实验室)
;
Academy of Military Science(军事科学院)
;
Bosch innovation software development (Wuxi) Co., Ltd.(博世创新软件开发(无锡)有限公司)
;
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机学院)
;
Shopee Pte. Ltd.(Shopee私人有限公司)
Predicting Immune Biomarkers with MultiModal Mixture-of-Expert Pathology Foundation Models Empowers Precision Oncology
使用多模态混合专家病理基础模型预测免疫生物标志物,赋能精准肿瘤学
Tianyu Liu, Ziqing Wang, Zhaokang Liang, Tong Ding, Peter Humphrey, Lorraine Colón-Cartagena, Emily Ling-Lin Pai, Kenneth Tou En Chang, Mohamed Kahila, Jonathan Chong Kai Liew, Tinglin Huang, Rex Ying, Kaize Ding, Faisal Mahmood, Wengong Jin
机构
*
Program of Computational Biology and Bioinforamtics, Yale University(耶鲁大学计算生物学与生物信息学项目)
;
Broad Institute of MIT and Harvard(麻省理工学院与哈佛大学博德研究所)
;
Department of Statistics and Data Science, Northwestern University(西北大学统计与数据科学系)
;
Department of Computer Science, Northeastern University(东北大学计算机科学系)
;
Department of Computer Science, Harvard University(哈佛大学计算机科学系)
;
Department of Pathology, Yale University(耶鲁大学病理学系)
;
Department of Anatomic Pathology and Laboratory Medicine, Hospital of the University of Pennsylvania(宾夕法尼亚大学医院解剖病理学与检验医学系)
;
Department of Pathology and Laboratory Medicine, University of California, San Francisco(加州大学旧金山分校病理学与检验医学系)
;
Department of Pathology and Laboratory Medicine, KK Women’s and Children’s Hospital(竹脚妇幼医院病理学与检验医学系)
;
Department of Biostatistics, Epidemiology and Informatics, Perelman School of Medicine, University of Pennsylvania(宾夕法尼亚大学佩雷尔曼医学院生物统计学、流行病学与信息学系)
专题命中
多模态生成
:multimodal(title,abstract);multimodal foundation model(abstract);分类 cs.CV
UniMedVL: Unifying Medical Multimodal Understanding and Generation through Observation-Knowledge-Analysis
UniMedVL: 通过观察-知识-分析统一医学多模态理解与生成
Junzhi Ning, Wei Li, Cheng Tang, Jiashi Lin, Chenglong Ma, Chaoyang Zhang, Jiyao Liu, Ying Chen, Shujian Gao, Yuandong Pu, Huihui Xu, Chenhui Gou, Ziyan Huang, Yi Xin, Qi Qin, Diping Song, Bin Fu, Guang Yang, Yuanfeng Ji, Tianbin Li, Yanzhou Su, Jin Ye, Shixiang Tang, Zhongying Deng, Lihao Liu, Ming Hu, Junjun He
机构
*
Shanghai Artificial Intelligence Laboratory
;
Shanghai Innovation Institute
;
Shanghai Jiao Tong University
;
Shanghai Institute of Optics
;
Fudan University
;
University of Cambridge
;
Monash University
;
DAMO Academy, Alibaba Group
;
Imperial College London
;
The University of Hong Kong
;
The Hong Kong University of Science
;
Hupan Lab
;
The Chinese University of Hong Kong
机构
*
Tencent Youtu Lab(腾讯优图实验室)
;
Tsinghua University(清华大学)
;
The University of Hong Kong(香港大学)
;
University of Warwick(沃林汉大学)
;
Monash University(墨尔本大学)
;
The Hong Kong Polytechnic University(香港理工大学)