AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge Graphs
AISE-Bench:用于学术知识图谱信息检索的全周期精选基准测试
Fanjin Zhang, Zhengyang Wang, Ruixuan Huang, Kefan Zhang, Amy Xin, Yuanchun Wang, Shu Zhao, Evgeny Kharlamov, Jie Tang, Juanzi Li
机构
*
Renmin University of China(中国人民大学)
;
Anhui University(安徽大学)
;
Z-Lab Z.ai
;
Tsinghua University(清华大学)
;
Bosch Center for AI(博世人工智能中心)
;
University of Oslo(奥斯陆大学)
Journal refProceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD '26), August 09-13, 2026, Jeju Island, Republic of Korea
Benchmarking Resource-Efficient LLMs for Research Topic Ontology Generation in the Biomedical Field
用于生物医学领域研究主题本体生成的资源高效语言模型基准测试
Tanay Aggarwal, Angelo Salatino, Francesco Osborne, Enrico Motta
机构
*
Knowledge Media Institute, The Open University, Milton Keynes, UK(开放大学知识媒体研究所)
;
Department of Business and Law, University of Milano-Bicocca, Milan, IT(米兰-比科卡大学商业与法律系)
Optimizing Large Language Models for Causality Assessment in Pharmacovigilance: Developing a Performance Metric as Objective for Bayesian Hyperparameter Optimization
优化用于药物警戒中因果关系评估的大语言模型:开发一种性能指标作为贝叶斯超参数优化的目标
Nicole Sonne Heckmann, Arnault-Quentin Vermillet, Søren Norlin Mølgaard, Manuela Del Castillo Suero, Lars Melskens, Gerard Ompad, Maurizio Sessa
机构
*
Department of Drug Design and Pharmacology, University of Copenhagen(哥本哈根大学药物设计与药理学系)
;
Safety Operations, Novo Nordisk(诺和制药安全运营部)
;
Clinical Development Centre Denmark, Novo Nordisk(丹麦临床开发中心,诺和制药)
;
Statistical Data Science Lead, Pfizer(辉瑞统计数据科学负责人)
Embodied-BenchClaw: An Autonomous Multi-Agent System for Embodied Spatial Intelligence Benchmark Construction
Embodied-BenchClaw:用于具身空间智能基准构建的自主多智能体系统
Baoyang Jiang, Fengchun Zhang, Leyuan Wang, Haotian Li, Yida Wang, Zhe Ji, Jinshan Lai, Xi Ren, Jianwei Hu, Qiang Ma
机构
*
QiYuan Lab(启元实验室)
;
School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院)
;
School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院)
How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figures
视觉语言模型(VLMs)在失明或被误导时的表现如何?针对科学图表的VLMs行为评估
Paul Osemudiame Oamen, Owusu-Banahene Osei, Ananya Mukherjee, Christian Greisinger, Steffen Eger, Pius Onobhayedo, Wei Zhao
机构
*
University of Aberdeen(阿伯丁大学)
;
International Institute of Information Technology Hyderabad(海德拉巴国际信息技术学院)
;
University of Technology Nuremberg(纽伦堡工业大学)
;
University of Southern California(南加利福尼亚大学)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Large Language Model Department, Tencent(腾讯大语言模型部)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Zhongguancun Academy(中关村学院)