Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval
超越语义检索:面向组合图像检索的指代锚定
Yuxin Yang, Yinan Zhou, Yuxin Chen, Ziqi Zhang, Zongyang Ma, Chunfeng Yuan, Bing Li, Jun Gao, Weiming Hu
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Tencent Inc.(腾讯公司)
;
PeopleAI Inc.(人民人工智能公司)
;
HelloGroup Inc.(哈啰集团)
;
ShanghaiTech University(上海科技大学)
TeamPath: Building MultiModal Pathology Experts with Reasoning AI Copilots
TeamPath: 构建多模态病理专家的推理AI助手
Tianyu Liu, Weihao Xuan, Hao Wu, Peter Humphrey, Marcello DiStasio, Mohamed Kahila, Alfonso Garcia Tan, Heli Qi, Rui Yang, Simeng Han, Tinglin Huang, Fang Wu, Chen Liu, Qingyu Chen, Nan Liu, Irene Li, Hua Xu, Hongyu Zhao
机构
*
Interdepartmental Program in Computational Biology and Biomedical Informatics, Yale University(耶鲁大学计算生物学与生物医学信息学跨学科项目)
;
Department of Biostatistics, Yale University(耶鲁大学生物统计学系)
;
Broad Institute of MIT and Harvard(博德研究所)
;
Department of Complexity Science and Engineering, The University of Tokyo(东京大学复杂科学与工程系)
;
Center for Advanced Intelligence Project, RIKEN(理化学研究所先进智能项目中心)
;
Department of Pathology, Yale University(耶鲁大学病理学系)
;
Department of Anatomical Pathology, Singapore General Hospital(新加坡中央医院解剖病理学系)
;
Center for Biomedical Data Science, Duke–NUS Medical School, Singapore, Singapore(杜克-新加坡国立大学医学院生物医学数据科学中心)
;
Department of Computer Science, Yale University(耶鲁大学计算机科学系)
;
Department of Computer Science, Stanford University(斯坦福大学计算机科学系)
机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身智能研究所)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai Key Laboratory of Multimodal Embodied AI(上海市多模态具身人工智能重点实验室)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
The Chinese University of Hong Kong(香港中文大学)
;
Central South University(中南大学)
;
Fudan University(复旦大学)
General Multimodal Protein Design Enables DNA-Encoding of Chemistry
通用多模态蛋白质设计使化学编码成为可能
Jarrid Rector-Brooks, Théophile Lambert, Marta Skreta, Daniel Roth, Yueming Long, Zi-Qi Li, Xi Zhang, Miruna Cretu, Francesca-Zhoufan Li, Tanvi Ganapathy, Emily Jin, Avishek Joey Bose, Jason Yang, Kirill Neklyudov, Yoshua Bengio, Alexander Tong, Frances H. Arnold, Cheng-Hao Liu
机构
*
California Institute of Technology(加州理工学院)
;
Mila – Québec AI Institute(Mila – 魁北克人工智能研究所)
;
Université de Montréal(蒙特利尔大学)
;
Université Paris-Saclay(巴黎-萨克雷大学)
;
McGill University(麦吉尔大学)
;
University of Cambridge(剑桥大学)
;
University of Oxford(牛津大学)
;
Imperial College London(伦敦帝国理工学院)
;
Institut Courtois(库尔图瓦研究所)
;
LawZero
;
AITHYRA
;
FutureHouse
ProMQA-Assembly: Multimodal Procedural QA Dataset on Assembly
ProMQA-Assembly:多模态装配任务问答数据集
Kimihiro Hasegawa, Wiradee Imrattanatrai, Masaki Asada, Susan Holm, Yuran Wang, Vincent Zhou, Ken Fukuda, Teruko Mitamura
机构
*
Language Technologies Institute, Carnegie Mellon University(卡内基梅隆大学语言技术研究所)
;
National Institute of Advanced Industrial Science and Technology (AIST)(国立研究开发法人产业技术综合研究所(AIST))
机构
*
School of Computer Science and Technology, Eastern Institute of Technology, Ningbo, China(东部理工学院计算机科学与技术学院)
;
School of Mechatronic Engineering and Automation, Shanghai University, China(上海大学机电工程与自动化学院)
Recent Advances in Multimodal Affective Computing: An NLP Perspective
多模态情感计算近期进展:自然语言处理视角
Guimin Hu, Weimin Lyu, Chang Sun, Zhihong Zhu, Lin Gui, Ruichu Cai, Erik Cambria, Hasti Seifi
机构
*
Guangdong University of Technology(广东工业大学)
;
Stony Brook University(石溪大学)
;
University of Bologna(博洛尼亚大学)
;
Peking University(北京大学)
;
Nanyang Technological University(南洋理工大学)
;
Arizona State University(亚利桑那州立大学)
CommentsAccepted at the IMAGE'25 Workshop (PCW-11), Society of Exploration Geophysicists (SEG). Published version available at https://doi.org/10.1190/image2025-w11-03.1
CommentsXIAOHE Medical AI team. See paper for full author list. Currently, the model is exclusively available on XIAOHE AI Doctor, accessible via both the App Store and the Douyin Mini Program. Updated to improve the layout
机构
*
Yale University(耶鲁大学)
;
Broad Institute of MIT and Harvard(麻省理工学院-哈佛大学博德研究所)
;
Google DeepMind(谷歌DeepMind)
;
Stanford University(斯坦福大学)
;
Genentech(基因泰克)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
Cornell University(康奈尔大学)
;
Harvard University(哈佛大学)
Integration of Object Detection and Small VLMs for Construction Safety Hazard Identification
对象检测与小型视觉语言模型的整合用于建筑安全危险识别
Muhammad Adil, Mehmood Ahmed, Muhammad Aqib, Vicente A. Gonzalez, Gaang Lee, Qipei Mei
机构
*
Infrastructure Human Tech (IHT) Lab, Department of Civil and Environmental Engineering, University of Alberta, Edmonton, Alberta, Canada(基础设施人类技术(IHT)实验室,土木与环境工程系,阿尔伯塔大学,埃德蒙顿,阿尔伯塔,加拿大)