Comments7 pages. Corrects the accent-correctness evaluation to the crowdsourced listening test (v3 inadvertently reported the earlier author-scored results), and adds a note on the precedence of the proposed token-based pronunciation-control method relative to a subsequent technical report, together with links to the released code, training/evaluation data, LoRA weights, and audio samples
Chuanhao Yan, Fengdi Che, Xuhan Huang, Xu Xu, Xin Li, Yizhi Li, Xingwei Qu, Jingzhe Shi, Chenghua Lin, Yaodong Yang, Binhang Yuan, Hang Zhao, Yu Qiao, Bowen Zhou, Jie Fu
机构
*
Shanghai AI Lab(上海人工智能实验室)
;
University of Alberta(阿尔伯塔大学)
;
Tsinghua University(清华大学)
;
Chinese University of Hong Kong, Shenzhen(香港大学(深圳))
;
Hong Kong University of Science and Technology(香港科技大学)
;
Nanyang Technological University(南洋理工大学)
;
University of Manchester(曼彻斯特大学)
;
Peking University(北京大学)
专题命中
指令微调
:large language model(abstract);language model(abstract);SFT(abstract);分类 cs.CL
机构
*
Carnegie Mellon University(卡内基梅隆大学)
;
Jinesis Lab, University of Toronto & Vector Institute(Jinesis实验室,多伦多大学及向量研究所)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Princeton University(普林斯顿大学)
;
Cornell University(康奈尔大学)
;
The University of Tokyo(东京大学)
;
RIKEN AIP(日本理化学研究所AIP)
;
Max Planck Institute for Intelligent Systems, Tübingen, Germany(德国图宾根最大计划智能系统研究所)
;
EuroSafeAI
Med-R2: Perception and Reflection-driven Complex Reasoning for Medical Report Generation
Med-R2:面向医学报告生成的感知与反思驱动复杂推理
Hao Wang, Shuchang Ye, Jinghao Lin, Usman Naseem, Jinman Kim
机构
*
The School of Computer Science, The University of Sydney(悉尼大学计算机科学学院)
;
The School of Computing, Macquarie University(麦考瑞大学计算机学院)
;
Doubao Medical Group, ByteDance(字节跳动 doubao 医疗集团)
Comments20 pages, 5 figures, 7 tables. Major revision and repositioning of arXiv:2504.15610v1-v3 (previously titled "A LoRA-Based Approach to Fine-Tuning LLMs for Educational Guidance in Resource-Constrained Settings"); withdraws the earlier quantization-boundary and cross-GPU optimizer-transfer claims. Code, dataset, adapter, and evaluation harness released
Operationalising the Superficial Alignment Hypothesis via Task Complexity
通过任务复杂度操作化浅层对齐假设
Tomás Vergara-Browne, Darshan Patil, Ivan Titov, Siva Reddy, Tiago Pimentel, Marius Mosbach
机构
*
University of Maryland(马里兰大学)
;
University of California, Berkeley(加州大学伯克利分校)
;
University of Washington(华盛顿大学)
;
University of Toronto(多伦多大学)
;
University of Edinburgh(爱丁堡大学)
专题命中
指令微调
:large language model(abstract);language model(abstract);post-training(abstract);分类 cs.LG
Role Steering of Language Models for Social Simulations
用于社会模拟的语言模型角色引导
Isaac Song, Mohammed Rehan Parwani, Glenn Matlin, Emile Anand, Akhil Theerthala, Arjun Chatterjee, Anthony Wen-Ming Zang, Maria Kostylew, Yonadav G. Shavit, Sebastien Krier, Mark Riedl
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
ML Alignment & Theory Scholars (MATS)(ML对齐与理论学者组织(MATS))
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Oxford(牛津大学)
;
Google DeepMind(谷歌DeepMind)
;
OpenAI(开放人工智能公司)
Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for Secure Network Electronic Health Record (EHR) Interoperability
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
预条件测试时适应用于叙事生成中的分布外去偏
Hanwen Shen, Ting Ying, Jiajie Lu, Shanshan Wang
机构
*
Laboratory for Artificial Intelligence in Mathematics Education, Stevens Institute of Technology(数学教育中的人工智能实验室,史蒂文斯理工学院)
;
Independent Researcher(独立研究者)
;
NLP2CT Lab, Department of Computer and Information Science, University of Macau(NLP2CT实验室,澳门大学计算机与信息科学系)
专题命中
指令微调
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI