机构
*
National Taiwan University(国立台湾大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Academia Sinica, Taiwan(台湾“中央研究院”)
;
NTU AI-CoRE(国立清华大学AI研究中心)
专题命中
评测与基准
:LLM(summary_cn,abstract_cn);large language model(title);language model(title);分类 cs.CL、cs.AI
机构
*
University of Science and Technology of China(中国科学技术大学)
;
The First Affiliated Hospital of USTC(中国科学技术大学附属第一医院)
;
National University of Singapore(新加坡国立大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
LLM4AD: Large Language Models for Autonomous Driving -- Concept, Review, Benchmark, Experiments, and Future Trends
LLM4AD:大型语言模型在自动驾驶中的应用——概念、综述、基准测试、实验与未来趋势
Can Cui, Yunsheng Ma, Sung-Yeon Park, Zichong Yang, Yupeng Zhou, Peiran Liu, Juanwu Lu, Juntong Peng, Jiaru Zhang, Ruqi Zhang, Lingxi Li, Yaobin Chen, Jitesh H. Panchal, Amr Abdelraouf, Rohit Gupta, Kyungtae Han, Ziran Wang
机构
*
Purdue University(普渡大学)
;
Institute for Physical Artificial Intelligence (IPAI), Purdue University(普渡大学物理人工智能研究所)
;
Department of Computer Science, Purdue University(普渡大学计算机科学系)
;
InfoTech Labs, Toyota Motor North America(丰田北美信息技术实验室)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
SQLBench: A Comprehensive Evaluation for Text-to-SQL Capabilities of Large Language Models
SQLBench: 一种全面评估大型语言模型文本到SQL能力的综合评估
Bin Zhang, Yuxiao Ye, Guoqing Du, Xiaoru Hu, Zhishuai Li, Chi Harold Liu, Zhiwei Xu, Guoliang Fan, Rui Zhao, Ziyue Li, Hangyu Mao
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
SenseTime Research(商汤科技研究院)
;
School of Artificial Intelligence, Shandong University(山东大学人工智能学院)
;
School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院)
;
Technical University of Munich, Heilbronn Data Science Center, Munich Data Science Institute(慕尼黑技术大学,海德堡数据科学中心,慕尼黑数据科学研究所)
;
Institute of Microelectronics, Chinese Academy of Sciences(中国科学院微电子研究所)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
How Large Language Models Get Stuck: Early structure with persistent errors
大语言模型为何陷入停滞:早期结构与持续性错误
Alokesh Manna, William Snyder, Whitney Tabor
机构
*
Department of Statistics(统计学系)
;
University of Connecticut, Storrs(康涅狄格大学斯托尔分校)
;
Department of Linguistics(语言学系)
;
Department of Psychological Sciences(心理学科学系)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.LG
Large Language Model Psychometrics: A Systematic Review of Evaluation, Validation, and Enhancement
大语言模型心理测量学:评估、验证与增强的系统综述
Haoran Ye, Jing Jin, Yuhang Xie, Xin Zhang, Guojie Song
机构
*
State Key Laboratory of General Artificial Intelligence, School of Intelligence Science and Technology, Peking University(一般人工智能国家重点实验室,智能科学与技术学院,北京大学)
;
School of Psychological and Cognitive Sciences, Peking University(心理学与认知科学学院,北京大学)
;
Key Laboratory of Machine Perception (Ministry of Education), Peking University(机器感知重点实验室(教育部),北京大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI
A Component-Based Survey of Interactions between Large Language Models and Multi-Armed Bandits
基于组件的大型语言模型与多臂老虎机交互调研
Siguang Chen, Chunli Lv, Miao Xie
机构
*
College of Information and Electrical Engineering, China Agricultural University, Beijing 100083, China(信息与电气工程学院,中国农业大学,北京)
;
Key Laboratory of Agricultural Machinery Monitoring and Big Data Application, Ministry of Agriculture and Rural Affairs, Beijing 100083, China(农业机械监测与大数据应用重点实验室,农业农村部,北京)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.LG