LLMs can be easily Confused by Instructional Distractions
Yerin Hwang, Yongil Kim, Jahyun Koo, Taegwan Kang, Hyunkyung Bae, Kyomin Jung
机构
*
IPAI, Seoul National University(IPAI,首尔国立大学)
;
LG AI Research(LG人工智能研究所)
;
Dept. of ECE, Seoul National University(电子工程系,首尔国立大学)
;
SNU-LG AI Research Center(SNU-LG人工智能研究所)
专题命中
数学推理
:reasoning(abstract);分类 cs.CL、cs.AI
Comments8 pages
Journal refProceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), pages 19483-19496, Vienna, Austria, July 2025
A Practical Two-Stage Recipe for Mathematical LLMs: Maximizing Accuracy with SFT and Efficiency with Reinforcement Learning
Hiroshi Yoshihara, Taiki Yamaguchi, Yuichi Inoue
机构
*
Aillis Inc.(Aillis公司)
;
Department of Health Policy(健康政策部门)
;
Public Health, Graduate School of Pharmaceutical Sciences, The University of Tokyo, Tokyo, Japan(公共卫生,药学研究生院,东京大学,东京,日本)
;
Rist Inc.(Rist公司)
专题命中
数学推理
:reasoning(abstract);分类 cs.AI、cs.LG
CommentsPresented at ICML 2025 Workshop on The second AI for MATH
LoRA-One: One-Step Full Gradient Could Suffice for Fine-Tuning Large Language Models, Provably and Efficiently
Yuanhe Zhang, Fanghui Liu, Yudong Chen
机构
*
Department of Statistics, University of Warwick, UK.(威斯敏斯特大学统计学系)
;
Department of Computer Science, University of Warwick, UK.(威斯敏斯特大学计算机科学系)
;
Centre for Discrete Mathematics and its Applications (DIMAP), University of Warwick, UK.(威斯敏斯特大学离散数学及其应用中心(DIMAP))
Zhiyuan Liang, Dongwen Tang, Yuhao Zhou, Xuanlei Zhao, Mingjia Shi, Wangbo Zhao, Zekai Li, Peihao Wang, Konstantin Schürholt, Damian Borth, Michael M. Bronstein, Yang You, Zhangyang Wang, Kai Wang
机构
*
National University of Singapore(新加坡国立大学)
;
UT Austin(德克萨斯大学奥斯汀分校)
;
University of St. Gallen(圣加尔登大学)
;
Oxford University(牛津大学)
专题命中
数学推理
:reasoning(abstract);分类 cs.AI、cs.LG
CommentsWe propose a method that can generate LoRA parameters in seconds
Investigating the Potential of Large Language Model-Based Router Multi-Agent Architectures for Foundation Design Automation: A Task Classification and Expert Selection Study
Sompote Youwai, David Phim, Vianne Gayl Murcia, Rianne Clair Onas
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Tsinghua University(清华大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(教育部下一代智能搜索与推荐工程技术研究中心)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
Wait, that's not an option: LLMs Robustness with Incorrect Multiple-Choice Options
Gracjan Góral, Emilia Wiśnios, Piotr Sankowski, Paweł Budzianowski
机构
*
University of Warsaw(华沙大学)
;
Institute of Mathematics, Polish Academy of Sciences(波兰科学院数学研究所)
;
MIM Solutions(MIM解决方案)
;
K-Scale Labs(K-Scale实验室)
;
IDEAS NCBR
;
IDEAS Research Institute(IDEAS研究学院)
专题命中
数学推理
:reasoning(abstract);分类 cs.CL、cs.AI
CommentsAccepted for ACL 2025 Main Conference and NeurIPS 2024 FM-EduAssess Workshop