arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Annual Meeting of the Association for Computational Linguistics · 会议 · Natural Language Processing

2026-03-31 至 2026-03-31 共收录 5
2603.27877 2026-03-31 cs.CL cs.SD

HumMusQA: A Human-written Music Understanding QA Benchmark Dataset

HumMusQA:一个人类撰写的音乐理解问答基准数据集

Benno Weck, Pablo Puentes, Andrea Poltronieri, Satyajeet Prabhu, Dmitry Bogdanov

机构 * Universitat Pompeu Fabra(庞培法布拉大学) Universitat Autònoma de Barcelona(巴塞罗那自治大学)

AI总结 本文提出HumMusQA数据集,通过专家手工构建320个问题,评估大型音频-语言模型对音乐的理解能力,并测试其对单模态捷径的鲁棒性。

Comments Dataset available at https://doi.org/10.5281/zenodo.18462523

Journal ref Proceedings of the 4th Workshop on NLP for Music and Audio (NLP4MusA 2026), pages 58-67, Rabat, Morocco. Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21652 2026-03-31 cs.CL

UnsafeChain: Enhancing Reasoning Model Safety via Hard Cases

UnsafeChain: 通过困难案例增强推理模型安全性

Raj Vardhan Tomar, Preslav Nakov, Yuxia Wang

机构 * Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学) Cluster Innovation Centre, University of Delhi(德里大学集群创新中心) INSAIT

AI总结 UnsafeChain通过构建包含困难提示的对齐数据集,提升模型安全性并保持推理能力,实验显示其在多个基准测试中表现优异。

Journal ref The Asian Federation of Natural Language Processing and The Association for Computational Linguistics 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12616 2026-03-31 cs.CL

Improving Chain-of-Thought Reasoning via Quasi-Symbolic Abstractions

通过准符号抽象改进链式推理

Leonardo Ranaldi, Marco Valentino, Andrè Freitas

机构 * Idiap Research Institute(Idiap 研究所) School of Informatics, University of Edinburgh(爱丁堡大学信息学院) School of Computer Science, University of Sheffield(谢菲尔德大学计算机科学学院) Department of Computer Science, University of Manchester(曼彻斯特大学计算机科学系) National Biomarker Centre (NBC), CRUK Manchester Institute(英国癌症研究中心曼彻斯特研究所国家生物标志物中心)

AI总结 本文提出QuaSAR方法,通过准符号解释引导LLM在更高抽象层面推理,提升小模型的推理能力,实验显示在自然语言和符号推理任务中准确率提升8%。

Journal ref 2025.acl-long.843

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.26710 2026-03-31 cs.IR cs.AI cs.CL cs.MA

Agentic AI for Human Resources: LLM-Driven Candidate Assessment

面向人力资源的代理AI:基于大语言模型的候选人评估

Kamer Ali Yuksel, Abdul Basit Anees, Ashraf Elneima, Sanjika Hewavitharana, Mohamed Al-Badrashiny, Hassan Sawaf

机构 * aiXplain, Inc.(aiXplain公司)

AI总结 本文提出一个模块化且可解释的框架,利用大语言模型自动化招聘中的候选人评估。系统整合多种来源生成结构化评估报告,采用LLM生成的评分标准和多代理架构进行细粒度评估,输出透明、可审计的评估报告和推荐。

Comments Published in 19th Conference of the European Chapter of the Association for Computational Linguistics (EACL 2026)

Journal ref 19th Conference of the European Chapter of the Association for Computational Linguistics (EACL 2026), Rabat, Morocco

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.26671 2026-03-31 cs.LG math.OC

Mitigating Forgetting in Continual Learning with Selective Gradient Projection

通过选择性梯度投影缓解持续学习中的遗忘

Anika Singh, Aayush Dhaulakhandi, Varun Chopade, Likhith Malipati, David Martinez, Kevin Zhu

机构 * Algoverse AI Research(Algoverse AI 研究)

AI总结 本文提出SFAO方法,通过余弦相似度和分层门控调节梯度方向,实现受控遗忘并平衡可塑性与稳定性,实验显示其在持续学习基准上表现优异,内存成本降低90%。

Comments 15 pages, 2 figures, Accepted to the Student Research Workshop at International Joint Conference on Natural Language Processing & Asia-Pacific Chapter of the Association for Computational Linguistics, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏