Domyn-Small: A European 10B Reasoning Language Model
Domyn-Small:一个欧洲的100亿参数推理语言模型
Simone Angarano, Francesco Bertolotti, Federico D'Ambrosio, Michele Resta, Alessandro Rognoni, Nicolò Ruggeri, Dario Salvati, Andrea Valenti, Alberto Veneri, Martin Cimmino
机构
*
Vellore Institute of Technology(维洛雷理工学院)
;
University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校)
;
Northwestern University(西北大学)
;
Yale University(耶鲁大学)
;
Algoverse AI Research(Algoverse AI研究)
Comments12 pages, 4 figures, 6 tables. Includes ablation study across Qwen2.5-7B-Instruct and Llama-3.1-8B-Instruct on 5 math reasoning benchmarks (GSM8K, MATH500, Minerva, AIME24, Gaokao2023). GPT-4.1 used for structured evaluation of reasoning quality
机构
*
University of Arizona, USA(亚利桑那大学)
;
Arizona State University, USA(亚利桑那州立大学)
;
Now at Google LLC, work done at Rice University(现就职于谷歌公司,曾就职于里士大学)
;
Clemson University, USA(克莱姆森大学)
;
Washington University in St. Louis, USA(圣路易斯华盛顿大学)
;
Halmstad University, Sweden(哈姆斯塔德大学)
;
Guangdong Institute of Intelligence Science and Technology, China(广东智能科学与技术研究院)
Self-Prompting Small Language Models for Privacy-Sensitive Clinical Information Extraction
面向隐私敏感的临床信息抽取的自提示小型语言模型
Yao-Shun Chuang, Tushti Mody, Uday Pratap Singh, Shirindokht Shiraz, Chun-Teh Lee, Ryan Brandon, Muhammad F Walji, Xiaoqian Jiang, Bunmi Tokede
机构
*
McWilliams School of Biomedical Informatics, The University of Texas Health Science Center at Houston(德克萨斯大学健康科学中心休斯顿分校麦克威廉斯生物医学信息学学院)
;
School of Public Health, The University of Texas Health Science Center at Houston(德克萨斯大学健康科学中心休斯顿分校公共卫生学院)
;
School of Dentistry, The University of Texas Health Science Center at Houston(德克萨斯大学健康科学中心休斯顿分校牙科学院)
;
Willamette Dental and Skourtes Institute(威廉特牙科与斯库尔特斯研究所)
From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning
从消费到反思:为稳定推理设计人-人工智能关系
Rikard Rosenbacke, Carl Rosenbacke, Victor Rosenbacke, Martin McKee
机构
*
Faculty of Medicine, Lund University(吕勒欧大学医学院)
;
Department of Economics, Lund University School of Economics and Management(吕勒欧大学经济学与管理学院经济系)
;
Department of Health Services Research and Policy, London School of Hygiene & Tropical Medicine(伦敦卫生与热带医学学院健康服务研究与政策系)
Self-Improving Tabular Language Models via Iterative Reward-Guided Post-Training
通过迭代奖励引导的后训练改进表格语言模型
Yunbo Long, Tejumade Afonja, Guangya Hao, Alexandra Brintrup, Mario Fritz
机构
*
Department of Engineering, University of Cambridge(剑桥大学工程系)
;
CISPA Helmholtz Center for Information Security, Saarbrücken, Germany(德国萨尔布吕肯信息安全中心)
;
The Alan Turing Institute, London(伦敦阿兰·图灵研究所)