arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Annual Meeting of the Association for Computational Linguistics · 会议 · Natural Language Processing

共收录 10306
2506.00264 2025-06-05 cs.CL

MultiHoax: A Dataset of Multi-hop False-Premise Questions

Mohammadamin Shafiei, Hamidreza Saffari, Nafise Sadat Moosavi

机构 * University of Milan(米兰大学) Politecnico di Milano(米兰理工学院) University of Sheffield(谢菲尔德大学)

Comments accepted at ACL Findings 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14175 2025-06-05 cs.CL cs.IR

Hypothetical Documents or Knowledge Leakage? Rethinking LLM-based Query Expansion

Yejun Yoon, Jaeyoon Jung, Seunghyun Yoon, Kunwoo Park

机构 * Department of Intelligent Semiconductors, Soongsil University(智能半导体系,顺斯尔大学) School of AI Convergence, Soongsil University(人工智能融合学院,顺斯尔大学) MAUM AI Inc.(MAUM人工智能公司) Adobe Research, USA(Adobe美国研究实验室)

Comments ACL 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23899 2025-06-05 cs.CL

Rubrik's Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset

Diana Galvan-Sosa, Gabrielle Gaudeau, Pride Kavumba, Yunmeng Li, Hongyi gu, Zheng Yuan, Keisuke Sakaguchi, Paula Buttery

机构 * ALTA Institute, Computer Laboratory, University of Cambridge(ALTA研究所、计算机实验室、剑桥大学) SB Intuitions Tohoku University(东北大学) RIKEN(日本理化学研究所) The University of Sheffield(谢菲尔德大学)

Comments 10 main pages (24 appendix pages), 9 figures, accepted to ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15358 2025-06-05 cs.CL cs.CV

SemEval-2025 Task 1: AdMIRe -- Advancing Multimodal Idiomaticity Representation

Thomas Pickard, Aline Villavicencio, Maggie Mi, Wei He, Dylan Phelps, Marco Idiart

机构 * University of Sheffield, UK(谢菲尔德大学) University of Exeter, UK(埃克塞特大学) Federal University of Rio Grande do Sul, Brazil(里约格朗德杜斯尔大学)

Comments Author accepted version; SemEval-2025 proceedings to appear at ACL 2025. This version corrects a typo in the results table

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10515 2025-06-05 cs.CL

Probing LLMs for Multilingual Discourse Generalization Through a Unified Label Set

Florian Eichin, Yang Janet Liu, Barbara Plank, Michael A. Hedderich

机构 * MaiNLP, Center for Information and Language Processing, LMU Munich(MaiNLP、信息与语言处理中心、慕尼黑大学) Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)

Comments 18 pages, 7 figures, 3 tables, code: https://github.com/mainlp/discourse_probes, camera-ready revision for ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10267 2025-06-05 cs.CL

An Expanded Massive Multilingual Dataset for High-Performance Language Technologies (HPLT)

Laurie Burchell, Ona de Gibert, Nikolay Arefyev, Mikko Aulamo, Marta Bañón, Pinzhen Chen, Mariia Fedorova, Liane Guillou, Barry Haddow, Jan Hajič, Jindřich Helcl, Erik Henriksson, Mateusz Klimaszewski, Ville Komulainen, Andrey Kutuzov, Joona Kytöniemi, Veronika Laippala, Petter Mæhlum, Bhavitvya Malik, Farrokh Mehryary, Vladislav Mikhailov, Nikita Moghe, Amanda Myntti, Dayyán O'Brien, Stephan Oepen, Proyag Pal, Jousia Piha, Sampo Pyysalo, Gema Ramírez-Sánchez, David Samuel, Pavel Stepachev, Jörg Tiedemann, Dušan Variš, Tereza Vojtěchová, Jaume Zaragoza-Bernabeu

机构 * University of Edinburgh(爱丁堡大学) University of Helsinki(赫尔辛基大学) University of Oslo(奥斯陆大学) Prompsit Language Engineering(Prompsit语言工程) Charles University(查尔斯大学) University of Turku(图尔库大学)

Comments ACL'2025 Main Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08042 2025-06-05 cs.CL

LSC-Eval: A General Framework to Evaluate Methods for Assessing Dimensions of Lexical Semantic Change Using LLM-Generated Synthetic Data

Naomi Baes, Raphaël Merx, Nick Haslam, Ekaterina Vylomova, Haim Dubossarsky

机构 * Melbourne School of Psychological Sciences, The University of Melbourne(墨尔本大学心理学科学学院) School of Computing and Information Systems, The University of Melbourne(墨尔本大学计算与信息学院) School of Electronic Engineering and Computer Science, Queen Mary University of London(伦敦女王学院电子工程与计算机科学学院) The Alan Turing Institute, London(伦敦艾伦·图灵研究所) Language Technology Lab, University of Cambridge(剑桥大学语言技术实验室)

Comments Accepted to ACL Findings (9-page long paper; 35 pages total including limitations, appendices and references)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06474 2025-06-05 cs.IR cs.AI cs.CL

ROGRAG: A Robustly Optimized GraphRAG Framework

Zhefan Wang, Huanjun Kong, Jie Ying, Wanli Ouyang, Nanqing Dong

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Shanghai Innovation Institute(上海创新研究院) Chinese University of Hong Kong(香港中文大学)

Comments ACL2025 demo track, 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03962 2025-06-05 cs.CL

On the Acquisition of Shared Grammatical Representations in Bilingual Language Models

Catherine Arnett, Tyler A. Chang, James A. Michaelov, Benjamin K. Bergen

Comments 9 pages, 5 figures. Accepted at ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02769 2025-06-05 cs.SD cs.CL cs.HC eess.AS

InSerter: Speech Instruction Following with Unsupervised Interleaved Pre-training

Dingdong Wang, Jin Xu, Ruihang Chu, Zhifang Guo, Xiong Wang, Jincenzi Wu, Dongchao Yang, Shengpeng Ji, Junyang Lin

机构 * The Chinese University of Hong Kong(香港中文大学) Alibaba Group(阿里巴巴集团)

Comments Accepted to ACL 2025; Data is available at: https://huggingface.co/datasets/ddwang2000/SpeechInstructBench

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14748 2025-06-05 cs.CL

Large Language Models Struggle to Describe the Haystack without Human Help: Human-in-the-loop Evaluation of Topic Models

Zongxia Li, Lorena Calvo-Bartolomé, Alexander Hoyle, Paiheng Xu, Alden Dima, Juan Francisco Fung, Jordan Boyd-Graber

Comments 22 Pages. LLM for Data Exploration and content analysis, Topic Models. 63rd Annual Meeting of the Association for Computational Linguistics (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14019 2025-06-05 cs.CL cs.AI cs.HC

Dehumanizing Machines: Mitigating Anthropomorphic Behaviors in Text Generation Systems

Myra Cheng, Su Lin Blodgett, Alicia DeVrio, Lisa Egede, Alexandra Olteanu

机构 * Stanford University(斯坦福大学) Microsoft Research(微软研究院) Carnegie Mellon University(卡内基梅隆大学)

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13946 2025-06-05 cs.CL cs.AI cs.CR

Why Safeguarded Ships Run Aground? Aligned Large Language Models' Safety Mechanisms Tend to Be Anchored in The Template Region

Chak Tou Leong, Qingyu Yin, Jian Wang, Wenjie Li

机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学) Zhejiang University(浙江大学)

Comments ACL 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11753 2025-06-05 cs.AI

HintsOfTruth: A Multimodal Checkworthiness Detection Dataset with Real and Synthetic Claims

Michiel van der Meer, Pavel Korshunov, Sébastien Marcel, Lonneke van der Plas

Comments Accepted at ACL2025 (main track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05855 2025-06-05 cs.CL

ConSim: Measuring Concept-Based Explanations' Effectiveness with Automated Simulatability

Antonin Poché, Alon Jacovi, Agustin Martin Picard, Victor Boutin, Fanny Jourdan

Journal ref ACL 2025, Jul 2025, Vienna (Austria), France

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.21006 2025-06-05 cs.CL cs.AI

Verbosity-Aware Rationale Reduction: Effective Reduction of Redundant Rationale via Principled Criteria

Joonwon Jang, Jaehee Kim, Wonbin Kweon, Seonghyeon Lee, Hwanjo Yu

机构 * POSTECH Seoul National University(首尔国立大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Kyungpook National University(庆北国立大学) LG AI Research(LG AI研究院)

Comments ACL 2025 FINDINGS

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05237 2025-06-05 cs.CL cs.CV

MAmmoTH-VL: Eliciting Multimodal Reasoning with Instruction Tuning at Scale

Jarvis Guo, Tuney Zheng, Yuelin Bai, Bo Li, Yubo Wang, King Zhu, Yizhi Li, Graham Neubig, Wenhu Chen, Xiang Yue

机构 * Carnegie Mellon University(卡内基梅隆大学) M-A-P Nanyang Technological University(南洋理工大学) University of Waterloo(滑铁卢大学) The University of Manchester(曼彻斯特大学)

Comments ACL 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04920 2025-06-05 cs.CL cs.AI cs.DB

Enabling LLM Knowledge Analysis via Extensive Materialization

Yujia Hu, Tuan-Phong Nguyen, Shrestha Ghosh, Simon Razniewski

机构 * ScaDS.AI & TU Dresden(ScaDS.AI及德累斯顿技术大学) Max Planck Institute for Informatics(马克斯·普朗克信息研究所) University of Tübingen(图宾根大学)

Comments 14 pages, 4 tables, 12 figures

Journal ref ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14309 2025-06-05 cs.CL cs.AI

LoGU: Long-form Generation with Uncertainty Expressions

Ruihan Yang, Caiqi Zhang, Zhisong Zhang, Xinting Huang, Sen Yang, Nigel Collier, Dong Yu, Deqing Yang

机构 * Fudan University(复旦大学) University of Cambridge(剑桥大学) Tencent AI Lab(腾讯AI实验室) The Chinese University of Hong Kong(香港中文大学)

Comments ACL 2025 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09300 2025-06-05 cs.CL cs.AI cs.LG

Nudging: Inference-time Alignment of LLMs via Guided Decoding

Yu Fei, Yasaman Razeghi, Sameer Singh

机构 * Department of Computer Science University of California Irvine(计算机科学系,加州大学伊文斯顿分校)

Comments Accepted to ACL 2025 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.11547 2025-06-05 cs.CL cs.AI

Small Language Models can Outperform Humans in Short Creative Writing: A Study Comparing SLMs with Humans and LLMs

Guillermo Marco, Luz Rello, Julio Gonzalo

机构 * UNED, Madrid, Spain(UNED, 马德里, 西班牙) IE University, Madrid, Spain(IE大学, 马德里, 西班牙)

Comments Accepted as Main Conference Paper at COLING 2025

Journal ref Proceedings of the 31st International Conference on Computational Linguistics (COLING 2025), pages 6552-6570, Abu Dhabi, UAE. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01119 2025-06-05 cs.CL cs.AI

Pron vs Prompt: Can Large Language Models already Challenge a World-Class Fiction Author at Creative Text Writing?

Guillermo Marco, Julio Gonzalo, Ramón del Castillo, María Teresa Mateo Girona

机构 * School of Computer Science, UNED, Madrid, Spain(计算机科学学院,UNED,马德里,西班牙) Faculty of Education, UCM, Madrid, Spain(教育学院,UCM,马德里,西班牙) Faculty of Philosophy, UNED, Madrid, Spain(哲学学院,UNED,马德里,西班牙)

Comments 9 pages 6 figures

Journal ref Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, pages 19654-19670, Miami, Florida, USA. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12784 2025-06-05 cs.CL

UBench: Benchmarking Uncertainty in Large Language Models with Multiple Choice Questions

Xunzhi Wang, Zhuowei Zhang, Gaonan Chen, Qiongyu Li, Bitong Luo, Zhixin Han, Haotian Wang, Zhiyu li, Hang Gao, Mengting Hu

机构 * College of Software, Nankai University(南开大学软件学院) Institute for Advanced Algorithms Research (Shanghai)(上海先进算法研究所) College of Artificial Intelligence, Tianjin University of Science and Technology(天津科技大学人工智能学院) Tianjin Key Laboratory of Software Experience and Human Computer Interaction(天津软件体验与人机交互重点实验室)

Comments accepted by ACL Findings (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03149 2025-06-04 cs.CL cs.AI cs.LG

Causal Estimation of Tokenisation Bias

Pietro Lesci, Clara Meister, Thomas Hofmann, Andreas Vlachos, Tiago Pimentel

机构 * University of Cambridge(剑桥大学) ETH Zürich(苏黎世联邦理工学院)

Comments Published as a conference paper at ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03101 2025-06-04 cs.CL

Beyond Text Compression: Evaluating Tokenizers Across Scales

Jonas F. Lotz, António V. Lopes, Stephan Peitz, Hendra Setiawan, Leonardo Emili

机构 * University of Copenhagen(哥本哈根大学) ROCKWOOL Foundation Research Unit(ROCKWOOL基金会研究单位) Apple(苹果公司)

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03090 2025-06-04 cs.CL

Literary Evidence Retrieval via Long-Context Language Models

Katherine Thai, Mohit Iyyer

机构 * UMass Amherst(马萨诸塞大学阿姆赫斯特分校) University of Maryland, College Park(马里兰大学学院公园分校)

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18460 2025-06-04 cs.CL cs.IR

DRAMA: Diverse Augmentation from Large Language Models to Smaller Dense Retrievers

Xueguang Ma, Xi Victoria Lin, Barlas Oguz, Jimmy Lin, Wen-tau Yih, Xilun Chen

机构 * FAIR at Meta(Meta 的 FAIR 研究院) University of Waterloo(滑铁卢大学)

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13172 2025-06-04 cs.CR cs.AI

Unveiling Privacy Risks in LLM Agent Memory

Bo Wang, Weiyi He, Shenglai Zeng, Zhen Xiang, Yue Xing, Jiliang Tang, Pengfei He

机构 * Michigan State University(密歇根州立大学) University of Georgia(佐治亚大学)

Comments ACL 2025 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11184 2025-06-04 cs.CL cs.AI cs.CV cs.MM

Can't See the Forest for the Trees: Benchmarking Multimodal Safety Awareness for Multimodal LLMs

Wenxuan Wang, Xiaoyuan Liu, Kuiyi Gao, Jen-tse Huang, Youliang Yuan, Pinjia He, Shuai Wang, Zhaopeng Tu

机构 * Renmin University of China(中国人民大学) Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Chinese University of Hong Kong(香港中文大学) Johns Hopkins University(约翰霍普金斯大学) Hong Kong University of Science and Technology(香港科技大学) Tencent(腾讯)

Comments Accepted by ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09416 2025-06-04 cs.CL

Rethinking Evaluation Metrics for Grammatical Error Correction: Why Use a Different Evaluation Process than Human?

Takumi Goto, Yusuke Sakai, Taro Watanabe

机构 * Nara Institute of Science and Technology(奈良科学技术大学)

Comments ACL 2025 (Main), 5 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏