MaterialFigBENCH: benchmark dataset with figures for evaluating college-level materials science problem-solving abilities of multimodal large language models
MaterialFigBENCH:用于评估多模态大语言模型在大学级材料科学问题解决能力的基准数据集
Michiko Yoshitake, Yuta Suzuki, Ryo Igarashi, Yoshitaka Ushiku, Keisuke Nagato
STONE Dataset: A Scalable Multi-Modal Surround-View 3D Traversability Dataset for Off-Road Robot Navigation
STONE数据集:一个可扩展的多模态周围视图3D可通行性数据集用于越野机器人导航
Konyul Park, Daehun Kim, Jiyong Oh, Seunghoon Yu, Junseo Park, Jaehyun Park, Hongjae Shin, Hyungchan Cho, Jungho Kim, Jun Won Choi
机构
*
Interdisciplinary Program in Artificial Intelligence, Seoul National University(人工智能交叉学科项目,首尔国立大学)
;
Department of Electrical and Computer Engineering, Seoul National University(电气与计算机工程系,首尔国立大学)
CoMMET: To What Extent Can LLMs Perform Theory of Mind Tasks?
CoMMET:大型语言模型在理论思维任务中能发挥多大作用?
Ruirui Chen, Weifeng Jiang, Chengwei Qin, Cheston Tan
机构
*
Agency for Science, Technology and Research (A*STAR)(科技研究局)
;
Nanyang Technological University(南洋理工大学)
;
Hong Kong University of Science and Technology (Guangzhou), China(香港科技大学(广州),中国)