arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Annual Meeting of the Association for Computational Linguistics · 会议 · Natural Language Processing

2026-07-31 至 2026-07-31 共收录 3
2605.00086 2026-07-31 cs.CL cs.AI 版本更新

NorBERTo: A ModernBERT Model Trained for Portuguese with 331 Billion Tokens Corpus

NorBERTo:一个基于ModernBERT架构的葡萄牙语现代编码器模型,训练于331十亿个token语料库

Enzo S. N. Silva, Pablo B. Costa, Raphael C. Vlasman, Rosimeire P. Costa, Henrique L. P. Silva, Lucas F. A. O. Pellicer, Guilherme Rinaldo, Renato A. Almeida, Darian S. R. Rabbani, Cinthya O. Oestreich, Vinicius F. Caridá

机构 * Itaú Unibanco ICTi

AI总结 NorBERTo基于ModernBERT架构,利用Aurora-PT语料库训练,具备长上下文支持和高效注意力机制,在葡萄牙语语义相似性、文本蕴含和分类任务中表现优异,达到最高F1和准确率。

Comments This article has already undergone formal submission, review, acceptance, and publication in the proceedings of PROPOR 2026: Proceedings of the 17th International Conference on Computational Processing of Portuguese, Vol. 1. The published version is available in the ACL Anthology at https://aclanthology.org/2026.propor-1.18/ 11 pages, 9 tables, 2 figures

Journal ref Proceedings of the 17th International Conference on Computational Processing of Portuguese (PROPOR 2026) - Vol. 1

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18046 2026-07-31 cs.CV cs.AI 版本更新

Enhancing Scene Transition Awareness in Video Generation via Post-Training

通过后训练增强视频生成中的场景转换意识

Hanwen Shen, Jiajie Lu, Yupeng Cao, Xiaonan Yang

机构 * Stevens Institute of Technology(史蒂文斯理工学院)

AI总结 本文提出TAV数据集,通过后训练提升视频生成中场景转换的理解能力,改善多场景生成效果并保持图像质量。

Journal ref Proceedings of the 14th International Joint Conference on Natural Language Processing and the 4th Conference of the Asia-Pacific Chapter of the Association for Computational Linguistics (2025) 706-721

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06599 2026-07-31 cs.CL cs.AI 版本更新

How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs

上下文如何塑造真相:LLMs中语句级真相表示的几何变换

Shivam Adarsh, Maria Maistro, Christina Lioma

机构 * University of Copenhagen(哥本哈根大学)

AI总结 研究LLMs中上下文如何改变真相向量,发现早期层正交、中层收敛,上下文增加向量幅度,大模型通过方向变化区分相关与无关上下文。

Comments ACL 2026 (Main)

详情

展开后加载摘要…

URL PDF HTML 收藏