Conference on Empirical Methods in Natural Language Processing · 会议 · Natural Language Processing
共收录 7868 篇
2509.048662025-09-08cs.CL
Memorization $\neq$ Understanding: Do Large Language Models Have the Ability of Scenario Cognition?
Boxiang Ma, Ru Li, Yuanlong Wang, Hongye Tan, Xiaoli Li
机构
*
School of Computer and Information Technology, Shanxi University(山西大学计算机与信息学院)
;
Information Systems Technology and Design, Singapore University of Technology and Design(新加坡科技设计大学信息系统技术与设计)
Social Bias in Multilingual Language Models: A Survey
Lance Calvin Lim Gamboa, Yue Feng, Mark Lee
机构
*
School of Computer Science, University of Birmingham(伯明翰大学计算机科学学院)
;
Department of Information Systems and Computer Science, Ateneo de Manila University(马尼拉大学信息系统与计算机科学系)
机构
*
University of Colorado Anschutz Medical Campus(科罗拉多大学安舒茨医疗校园)
;
University of Colorado Boulder(科罗拉多大学博尔德分校)
;
University of Wisconsin Madison(威斯康星大学麦迪逊分校)
Conversational Education at Scale: A Multi-LLM Agent Workflow for Procedural Learning and Pedagogic Quality Assessment
Jiahuan Pei, Fanghua Ye, Xin Sun, Wentao Deng, Koen Hindriks, Junxiao Wang
机构
*
Vrije University of Amsterdam(阿姆斯特丹自由大学)
;
University College London(伦敦大学学院)
;
University of Amsterdam(阿姆斯特丹大学)
;
National Institute of Informatics(日本信息处理学会)
;
Shandong University(山东大学)
;
Guangzhou University(广州大学)
机构
*
University of Maryland, College Park(马里兰大学 College Park分校)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
;
University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校)
;
Georgia Tech(佐治亚理工学院)
;
Princeton University(普林斯顿大学)
Promptception: How Sensitive Are Large Multimodal Models to Prompts?
Mohamed Insaf Ismithdeen, Muhammad Uzair Khattak, Salman Khan
机构
*
Mohamed Bin Zayed University of Artificial Intelligence(莫罕默德·本·扎耶德人工智能大学)
;
Swiss Federal Institute of Technology Lausanne (EPFL)(洛桑联邦理工学院)
;
Australian National University(澳大利亚国立大学)
CANDY: Benchmarking LLMs' Limitations and Assistive Potential in Chinese Misinformation Fact-Checking
Ruiling Guo, Xinwei Yang, Chen Huang, Tong Zhang, Yong Hu
机构
*
School of Cyber Science and Engineering, Sichuan University, China(四川大学网络科学与工程学院)
;
College of Computer Science, Sichuan University, China(四川大学计算机学院)
;
Institute of Data Science, National University of Singapore, Singapore(新加坡国立大学数据科学研究所)
Forewarned is Forearmed: Pre-Synthesizing Jailbreak-like Instructions to Enhance LLM Safety Guardrail to Potential Attacks
Sheng Liu, Qiang Sheng, Danding Wang, Yang Li, Guang Yang, Juan Cao
机构
*
Sheng Liu Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences University of Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院,中国科学院大学)
;
Qiang Sheng Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院)
;
Danding Wang Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院)
;
Yang Li Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences University of Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院,中国科学院大学)
;
Guang Yang Zhongguancun Laboratory(中关村实验室)
;
Juan Cao Media Synthesis and Forensics Lab, Institute of Computing Technology, Chinese Academy of Sciences(媒体合成与取证实验室,计算技术研究所,中国科学院)
SLM-Bench: A Comprehensive Benchmark of Small Language Models on Environmental Impacts--Extended Version
Nghiem Thanh Pham, Tung Kieu, Duc-Manh Nguyen, Son Ha Xuan, Nghia Duong-Trung, Danh Le-Phuoc
机构
*
FPT University(FPT大学)
;
Aalborg University(奥尔堡大学)
;
Technische Universität Berlin(柏林技术大学)
;
RMIT University(皇家理工大学)
;
German Research Center for Artificial Intelligence(德国人工智能研究中心)
;
HiveIntel GmbH(HiveIntel公司)
Comments24 pages. An extended version of "SLM-Bench: A Comprehensive Benchmark of Small Language Models on Environmental Impacts" accepted at EMNLP 2025
Autoformalization in the Wild: Assessing LLMs on Real-World Mathematical Definitions
Lan Zhang, Marco Valentino, Andre Freitas
机构
*
Department of Computer Science, University of Manchester(曼彻斯特大学计算机科学系)
;
School of Computer Science, University of Sheffield(谢菲尔德大学计算机科学学院)
;
Idiap Research Institute(Idiap研究机构)
;
National Biomarker Centre, CRUK Manchester Institute(国家生物标志物中心、CRUK曼彻斯特研究所)