Imaging-101: Benchmarking LLM Coding Agents on Scientific Computational Imaging
成像101:在科学计算成像上对大语言模型编码智能体进行基准测试
Siyi Chen, Jiahe Ying, Yixuan Jia, Yuxuan Gu, Enze Ye, Weimin Bai, Zhijun Zeng, Shaochi Ren, Binhong Gao, Yubing Li, Tianhan Zhang, He Sun
机构
*
College of Future Technology and the National Biomedical Imaging Center, Peking University(北京大学未来技术学院和国家生物医学成像中心)
;
AI for Science Institute (AISI)(科学人工智能研究所)
;
University of Michigan(密歇根大学)
;
State Key Laboratory of Acoustics and Marine Information, Institute of Acoustics, Chinese Academy of Sciences(中国科学院声学研究所声场声信息国家重点实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
School of Astronautics, Beihang University(北京航空航天大学宇航学院)
;
Key Laboratory of Spacecraft Design Optimization and Dynamic Simulation Technologies, Ministry of Education(教育部航天器设计优化与动态仿真技术重点实验室)
机构
*
HKUST(GZ)(香港科技大学(广州))
;
Paradoox AI(悖论人工智能公司)
;
E Fund Management Co., Ltd(易方达基金管理有限公司)
;
MBZUAI(Mohamed Bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学))
;
The University of Tokyo(东京大学)
Evolution of Accuracy and Visual-Cognitive Errors in a Decade of Vision-Language AI Models
十年视觉语言人工智能模型中准确性和视觉认知错误的演变
Shravan Murlidaran, Miguel P. Eckstein
机构
*
Psychological & Brain Sciences, University of California, Santa Barbara(加利福尼亚大学圣巴巴拉分校心理与脑科学系)
;
Department of Computer Science, University of California, Santa Barbara(加利福尼亚大学圣巴巴拉分校计算机科学系)
;
Department of Electrical and Computer Engineering, University of California, Santa Barbara(加利福尼亚大学圣巴巴拉分校电气与计算机工程系)
A safety-oriented hypothetico-deductive framework for AI-assisted differential diagnosis
一种用于人工智能辅助鉴别诊断的面向安全的假设-演绎框架
Fan Ma, Mauro Giuffrè, Donald Wright, Kent McCann, Mark Iscoe, Lingfei Qian, Mingyang Jiang, Chi Wing Ng, Na Hong, Huan He, Cathy Shyr, Qingyu Chen, Lee Schwamm, Lucila Ohno-Machado, Hua Xu
机构
*
Yale School of Medicine, Yale University(耶鲁大学医学院,耶鲁大学)
;
Università degli Studi di Trieste(的里雅斯特大学)
;
Vanderbilt University(范德堡大学)
A knowledge-augmented dataset of high-risk driving scenarios with LLM annotations for autonomous driving
一个用于自动驾驶的具有大语言模型注释的高风险驾驶场景知识增强数据集
Heye Huang, Jingguang Li, Zhiyuan Zhou, Paul Liang, Mingyu Wu, Kitae Jang, Jianqiang Wang
机构
*
Korea Advanced Institute of Science and Technology(韩国科学技术院)
;
Fudan University(复旦大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Shanghai Jiao Tong University(上海交通大学)
;
Tsinghua University(清华大学)