机构
*
Zachry Department of Civil
;
Environmental Engineering, Texas A\&M University
;
Department of Computer Science
;
Engineering, Texas A\&M University
Evidence-based diagnostic reasoning with multi-agent copilot for human pathology
基于多智能体助手的证据驱动诊断推理
Luca L. Weishaupt, Chengkuan Chen, Drew F. K. Williamson, Richard J. Chen, Guillaume Jaume, Tong Ding, Bowen Chen, Anurag Vaidya, Long Phi Le, Guillaume Jaume, Ming Y. Lu, Faisal Mahmood
机构
*
Health Sciences and Technology, Harvard-MIT(哈佛-MIT健康科学与技术)
;
Department of Pathology, Massachusetts General Hospital, Harvard Medical School(麻省总医院病理科,哈佛医学院)
;
Cancer Program, Broad Institute of Harvard and MIT(哈佛-MIT博德研究所癌症项目)
;
Harvard John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保尔森工程与应用科学学院)
;
Electrical Engineering and Computer Science, Massachusetts Institute of Technology (MIT)(麻省理工学院电气工程与计算机科学)
;
Harvard Data Science Initiative, Harvard University(哈佛大学数据科学计划)
ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners?
ViGoR-Bench:视觉生成模型距离零样本视觉推理还有多远?
Haonan Han, Jiancheng Huang, Xiaopeng Sun, Junyan He, Rui Yang, Jie Hu, Xiaojiang Peng, Lin Ma, Xiaoming Wei, Xiu Li
机构
*
Tsinghua University(清华大学)
;
The University of Hong Kong(香港大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
IVEBench: Modern Benchmark Suite for Instruction-Guided Video Editing Assessment
IVEBench:现代指令引导视频编辑评估基准套件
Yinan Chen, Jiangning Zhang, Teng Hu, Yuxiang Zeng, Zhucun Xue, Qingdong He, Chengjie Wang, Yong Liu, Xiaobin Hu, Shuicheng Yan
机构
*
Zhejiang University(浙江大学)
;
Tencent Youtu Lab(腾讯优图实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Auckland(奥克兰大学)
;
National University of Singapore(新加坡国立大学)
Out-of-Sight Embodied Agents: Multimodal Tracking, Sensor Fusion, and Trajectory Forecasting
视线外的具身智能体:多模态跟踪、传感器融合与轨迹预测
Haichao Zhang, Yi Xu, Yun Fu
机构
*
Department of Electrical and Computer Engineering, Northeastern University(东北大学电气与计算机工程系)
;
Khoury College of Computer Sciences, Northeastern University(东北大学库里计算机科学学院)
Can a Robot Walk the Robotic Dog: Triple-Zero Collaborative Navigation for Heterogeneous Multi-Agent Systems
机器人能否行走:异构多智能体系统的三零协同导航
Yaxuan Wang, Yifan Xiang, Ke Li, Xun Zhang, BoWen Ye, Zhuochen Fan, Fei Wei, Tong Yang
机构
*
Yuanpei College, Peking University(北京大学元培学院)
;
School of Computer Science, Peking University(北京大学计算机学院)
;
School of Computer Science, Beijing University of Posts and Telecommunications(北京邮电大学计算机学院)
;
Pengcheng Laboratory(鹏城实验室)
;
Beijing Jinruyi Large Model Technology Co., Ltd.(北京金如意大模型科技有限公司)
CoVFT: Context-aware Visual Fine-tuning for Multimodal Large Language Models
CoVFT:面向多模态大语言模型的上下文感知视觉微调
Nan Zhou, Huiqun Wang, Yaoyan Zheng, Di Huang
机构
*
State Key Laboratory of Complex and Critical Software Environment, Beihang University(北京航空航天大学复杂关键软件环境国家重点实验室)
;
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能全国重点实验室)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Pecking University(北京大学)
;
Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所)
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Institute of Electronics and Information Industry Technology of Kashgar(喀什电子信息产业技术研究院)