机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身智能研究所)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai Key Laboratory of Multimodal Embodied AI(上海市多模态具身人工智能重点实验室)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
The Chinese University of Hong Kong(香港中文大学)
;
Central South University(中南大学)
;
Fudan University(复旦大学)
TeamPath: Building MultiModal Pathology Experts with Reasoning AI Copilots
TeamPath: 构建多模态病理专家的推理AI助手
Tianyu Liu, Weihao Xuan, Hao Wu, Peter Humphrey, Marcello DiStasio, Mohamed Kahila, Alfonso Garcia Tan, Heli Qi, Rui Yang, Simeng Han, Tinglin Huang, Fang Wu, Chen Liu, Qingyu Chen, Nan Liu, Irene Li, Hua Xu, Hongyu Zhao
机构
*
Interdepartmental Program in Computational Biology and Biomedical Informatics, Yale University(耶鲁大学计算生物学与生物医学信息学跨学科项目)
;
Department of Biostatistics, Yale University(耶鲁大学生物统计学系)
;
Broad Institute of MIT and Harvard(博德研究所)
;
Department of Complexity Science and Engineering, The University of Tokyo(东京大学复杂科学与工程系)
;
Center for Advanced Intelligence Project, RIKEN(理化学研究所先进智能项目中心)
;
Department of Pathology, Yale University(耶鲁大学病理学系)
;
Department of Anatomical Pathology, Singapore General Hospital(新加坡中央医院解剖病理学系)
;
Center for Biomedical Data Science, Duke–NUS Medical School, Singapore, Singapore(杜克-新加坡国立大学医学院生物医学数据科学中心)
;
Department of Computer Science, Yale University(耶鲁大学计算机科学系)
;
Department of Computer Science, Stanford University(斯坦福大学计算机科学系)
CommentsAccepted at the IMAGE'25 Workshop (PCW-11), Society of Exploration Geophysicists (SEG). Published version available at https://doi.org/10.1190/image2025-w11-03.1
Watch Before You Answer: Learning from Visually Grounded Post-Training
在回答前观看:从视觉引导的后训练中学习
Yuxuan Zhang, EunJeong Hwang, Huaisong Zhang, Penghui Du, Yiming Jia, Dongfu Jiang, Xuan He, Shenhui Zhang, Ping Nie, Peter West, Kelsey R. Allen
机构
*
University of British Columbia(不列颠哥伦比亚大学)
;
Vector Institute(向量研究所)
;
Etude AI
;
Kolors Team, Kuaishou Technology(快手科技Kolors团队)
;
University of Toronto(多伦多大学)
;
University of Waterloo(滑铁卢大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
CommentsXIAOHE Medical AI team. See paper for full author list. Currently, the model is exclusively available on XIAOHE AI Doctor, accessible via both the App Store and the Douyin Mini Program. Updated to improve the layout
Generating Synthetic Doctor-Patient Conversations for Long-form Audio Summarization
生成用于长文本音频摘要的合成医生-患者对话
Yanis Labrak, David Grünert, Séverin Baroudi, Jiyun Chun, Pawel Cyrta, Sergio Burdisso, Ahmed Hassoon, David Liu, Adam Rothschild, Reed Van Deusen, Petr Motlicek, Andrew Perrault, Ricard Marxer, Thomas Schaaf
机构
*
Idiap Research Institute(Idiap 研究所)
;
University of Zurich(苏黎世大学)
;
The Ohio State University(俄亥俄州立大学)
;
Université de Toulon, Aix Marseille Univ, LIS, CNRS(土伦大学,艾克斯-马赛大学,计算机与系统实验室,法国国家科学研究中心)
;
Stenograf
;
Johns Hopkins University Bloomberg School of Public Health(约翰·霍普金斯大学布隆伯格公共卫生学院)
;
Colorado School of Mines(科罗拉多矿业学院)
;
Allegheny Health Network(阿勒格尼健康网络)
;
University of Pittsburgh Medical Center(匹兹堡大学医学中心)
;
ILLS, CNRS(ILLS,法国国家科学研究中心)
;
Solventum
;
Carnegie Mellon University(卡内基梅隆大学)
Joint Knowledge Base Completion and Question Answering by Combining Large Language Models and Small Language Models
通过结合大语言模型和小语言模型实现知识库补全与问答的联合处理
Yinan Liu, Dongying Lin, Sigang Luo, Xiaochun Yang, Bin Wang
机构
*
School of Computer Science and Engineering, Northeastern University, Shenyang, China(东北大学计算机科学与工程学院,沈阳,中国)
;
National Frontiers Science Center for Industrial Intelligence and Systems optimization, Northeastern University, Shenyang, China(东北大学工业智能与系统优化国家级前沿科学中心,沈阳,中国)
From Human-Level AI Tales to AI Leveling Human Scales
从人类水平AI故事到AI人类尺度
Peter Romero, Fernando Martínez-Plumed, Zachary R. Tidler, Matthieu Téhénan, Sipeng Chen, Álvaro David Gómez Antón, Luning Sun, Manuel Cebrian, Lexin Zhou, Yael Moros Daval, Daniel Romero-Alvarado, Félix Martí Pérez, Kevin Wei, José Hernández-Orallo
机构
*
Valencian Research Institute of Artificial Intelligence, Universitat Politècnica de València(瓦伦西亚人工智能研究所,瓦伦西亚理工大学)
;
Leverhulme Centre for the Future of Intelligence, University of Cambridge(勒弗休姆未来智能中心,剑桥大学)
;
The Psychometrics Centre, University of Cambridge(心理测量中心,剑桥大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Harvard University(哈佛大学)
;
Center for Automation and Robotics, Spanish National Research Council(自动化与机器人中心,西班牙国家研究委员会)
;
Department of Computer Science, Princeton University(普林斯顿大学计算机科学系)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Cambridge, Department of Computer Sciences and Technology(剑桥大学计算机科学与技术系)
ProMQA-Assembly: Multimodal Procedural QA Dataset on Assembly
ProMQA-Assembly:多模态装配任务问答数据集
Kimihiro Hasegawa, Wiradee Imrattanatrai, Masaki Asada, Susan Holm, Yuran Wang, Vincent Zhou, Ken Fukuda, Teruko Mitamura
机构
*
Language Technologies Institute, Carnegie Mellon University(卡内基梅隆大学语言技术研究所)
;
National Institute of Advanced Industrial Science and Technology (AIST)(国立研究开发法人产业技术综合研究所(AIST))