机构
*
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
;
Zhongguancun Academy(中关村学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Nanyang Technological University(南洋理工大学)
;
Department of Mechanical Engineering, Imperial College London(伦敦帝国理工学院机械工程系)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);分类 cs.AI
Foundation Models for Credit Risk Prediction: A Game Changer?
信贷风险预测的基础模型:变革性突破?
Bart Baesens, Andreas Goethals, Stefan Lessmann, Simon De Vos, Cristián Bravo, David Martens, Victor Medina-Olivares, Christophe Mues, Maria Oskarsdóttir, Seppe vanden Broucke, Tony Van Gestel, Tim Verdonck, Wouter Verbeke
机构
*
Faculty of Economics and Business, KU Leuven, Belgium(比利时库勒万大学经济与商业学院)
;
School of Business and Economics, Humboldt University of Berlin, Germany(德国洪堡大学商学院)
;
Department of Statistical and Actuarial Sciences, Western University, Canada(加拿大西部大学统计与精算科学系)
;
Department of Engineering Management, University of Antwerp, Belgium(比利时安特卫普大学工程管理系)
;
Business School, University of Edinburgh, United Kingdom(英国爱丁堡大学商学院)
;
Business School, University of Southampton, United Kingdom(英国南安普顿大学商学院)
;
School of Mathematical Sciences, University of Southampton, United Kingdom(英国南安普顿大学数学科学学院)
;
Department of Business Informatics and Operations Management, Ghent University, Belgium(比利时根特大学商业信息与运营管理系)
;
Department of Mathematics, University of Antwerp, Belgium(比利时安特卫普大学数学系)
;
Department of Mathematics, KU Leuven, Belgium(比利时库勒万大学数学系)
专题命中
评测与基准
:foundation model(title,abstract);large language model(abstract);language model(abstract);pretraining(abstract)
Ask Before You Diagnose: Safe-Psych, a Sequential Evaluation Benchmark for LLMs in Psychiatry
诊断前先询问:Safe-Psych,一种用于精神病学领域大语言模型的顺序评估基准
Oriana Presacan, Andreea Grama, Larisa Irimină, Alireza Nik, Jaya Ojha, Vajira Thambawita, Ciprian I. Băcilă, Bogdan Ionescu, Michael A. Riegler
机构
*
National University of Science and Technology Politehnica Bucharest(布加勒斯特理工大学国立科技大学)
;
Psychiatric Hospital Doctor Gheorghe Preda(格奥尔基·普雷达医生精神病医院)
;
Oslo Metropolitan University(奥斯陆都市大学)
;
Kristiania University of Applied Sciences(克里斯蒂亚尼亚应用科学大学)
;
SimulaMet(SimulaMet公司)
;
Lucian Blaga University of Sibiu(锡比乌卢西安·布拉加大学)
专题命中
评测与基准
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);prompting(abstract)
LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents
LessonBench-V1:用于评估人工智能课程生成代理的基准数据集
Ravidu Suien Rammuni Silva, Ahmad Lotfi, Isibor Kennedy Ihianle, Golnaz Shahtahmassebi, Jordan J. Bird
机构
*
Department of Computer Science, Nottingham Trent University, Nottingham, UK(计算机科学系,诺丁汉特伦特大学,诺丁汉,英国)
;
Department of Physics and Mathematics, Nottingham Trent University, Nottingham, UK(物理与数学系,诺丁汉特伦特大学,诺丁汉,英国)
专题命中
评测与基准
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
CommentsTo appear at ICSE 2026. 13 pages. The L-AVRBench benchmark, Docker images, and evaluation scripts are available at https://github.com/rimwoohan/L-AVRBench