LINGOLY-TOO: Disentangling Reasoning from Knowledge with Templatised Orthographic Obfuscation
LINGOLY-TOO:通过模板化正交混淆分离推理与知识
Jude Khouja, Lingyi Yang, Karolina Korgul, Simeon Hellsten, Vlad A. Neacsu, Harry Mayne, Ryan Othniel Kearns, Andrew M. Bean, Adam Mahdi
机构
*
University of Oxford(牛津大学)
;
University of Nottingham(诺丁汉大学)
;
University of Glasgow(格拉斯哥大学)
;
United Kingdom Linguistics Olympiad(英国语言学奥林匹克)
;
Asia-Pacific Linguistics Olympiad(亚太语言学奥林匹克)
Improving Inverse Folding for Peptide Design with Diversity-regularized Direct Preference Optimization
通过多样性正则化的直接偏好优化改进肽设计的反向折叠
Ryan Park, Darren J. Hsu, C. Brian Roland, Maria Korshunova, Chen Tessler, Shie Mannor, Olivia Viessmann, Bruno Trentini
机构
*
Stanford University(斯坦福大学)
;
NVIDIA(英伟达)
;
Technion - Israel Institute of Technology(技术学院-以色列理工学院)
;
Flagship Pioneering(旗领先锋)
;
University of Oxford(牛津大学)
A Spectral Approach for Learning Spatiotemporal Neural Differential Equations
一种用于学习时空神经微分方程的谱方法
Mingtao Xia, Xiangting Li, Qijing Shen, Tom Chou
机构
*
Courant Institute of Mathematical Sciences, New York University, New York, NY 10012, USA(数学科学学院,纽约大学,纽约,NY 10012,美国)
;
Department of Computational Medicine, UCLA, Los Angeles, CA 90095, USA(计算医学系,洛杉矶大学,洛杉矶,CA 90095,美国)
;
Nuffield Department of Medicine, Oxford University, Oxford OX2 6HW, UK(医学系,牛津大学,牛津 OX2 6HW,英国)
Reasoning Compression with Mixed-Policy Distillation
基于混合策略蒸馏的推理压缩
Han Yang, Mingyan Wu, Bailan He, Zeyu Cao, Sikuan Yan, Kevin Qinghong Lin, Zifeng Ding
机构
*
Technical University of Munich(慕尼黑技术大学)
;
GESIS – Leibniz Institute for the Social Sciences(莱布尼茨社会科学研究所)
;
Northeastern University(东北大学)
;
LMU Munich(慕尼黑大学)
;
Siemens AG(西门子股份公司)
;
University of Cambridge(剑桥大学)
;
University of Oxford(牛津大学)
;
Mina AI
Quantile-Coupled Flow Matching for Distributional Reinforcement Learning
分位数耦合流匹配用于分布性强化学习
Michael Groom, Victor-Alexandru Darvariu, Lars Kunze, James Wilson, Nick Hawes
机构
*
Oxford Robotics Institute, University of Oxford(牛津大学机器人研究所)
;
Bristol Robotics Laboratory, University of the West of England(布里斯托尔机器人实验室,西英格兰大学)
;
Dyson Institute of Engineering & Technology(戴森工程与技术研究院)
Timo Stoll, Chendi Qian, Ben Finkelshtein, Ali Parviz, Darius Weber, Fabrizio Frasca, Hadar Shavit, Antoine Siraudin, Arman Mielke, Marie Anastacio, Erik Müller, Maya Bechler-Speicher, Michael Bronstein, Mikhail Galkin, Holger Hoos, Mathias Niepert, Bryan Perozzi, Jan Tönshoff, Christopher Morris
机构
*
RWTH Aachen University(亚琛RWTH大学)
;
University of Oxford(牛津大学)
;
Mila – Quebec AI Institute(魁北克AI研究所)
;
Technion - Israel Institute of Technology(技术学院-以色列理工学院)
;
ETAS Research University of Stuttgart(斯图加特大学ETAS研究所)
;
Google Research(谷歌研究)
;
Microsoft Research(微软研究院)
Compute as Teacher: Turning Inference Compute Into Reference-Free Supervision
计算作为教师:将推理计算转化为无参考监督
Dulhan Jayalath, Shashwat Goel, Thomas Foster, Parag Jain, Suchin Gururangan, Cheng Zhang, Anirudh Goyal, Alan Schelten
机构
*
University of Oxford(牛津大学)
;
ELLIS Institute Tübingen(图宾根ELLIS研究所)
;
Max Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所)
;
Meta Superintelligence Labs(元宇宙超级智能实验室)
OrderFusion: Encoding Orderbook for End-to-End Probabilistic Intraday Electricity Price Forecasting
OrderFusion:用于端到端概率日内电价预测的订单簿编码
Runyao Yu, Yuchen Tao, Fabian Leimgruber, Tara Esterl, Jochen Stiasny, Derek W. Bunn, Qingsong Wen, Hongye Guo, Jochen L. Cremer
机构
*
Delft University of Technology(代尔夫特理工大学)
;
Austrian Institute of Technology(奥地利技术研究院)
;
RWTH Aachen(亚琛工业大学)
;
London Business School(伦敦商学院)
;
University of Oxford(牛津大学)
;
Squirrel AI(Squirrel AI公司)
;
Tsinghua University(清华大学)
The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play
镜中的攻击者:通过锚定双策略自我博弈打破安全性中的自我一致性
Gabriele La Malfa, Emanuele La Malfa, Saar Cohen, Jie M. Zhang, Michael Luck, Michael Wooldridge, Elizabeth Black
机构
*
Department of Informatics, King’s College London(伦敦国王学院信息学院)
;
Department of Computer Science, University of Oxford(牛津大学计算机科学系)
;
University of Sussex(Sussex大学)
;
Institute for Decentralized AI(去中心化人工智能研究所)
Continuum Robot Localization using Distributed Time-of-Flight Sensors
基于分布式飞行时间传感器的连续机器人定位
Spencer Teetaert, Giammarco Caroleo, Marco Pontin, Sven Lilge, Jessica Burgner-Kahrs, Timothy D. Barfoot, Perla Maiolino
机构
*
University of Toronto Robotics Institute, Canada(多伦多大学机器人研究所,加拿大)
;
Oxford Robotics Institute, University of Oxford, UK(牛津大学机器人研究所,英国)
;
Toronto Metropolitan University, Canada(多伦多 Metropolitan 大学,加拿大)
Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners
玩乐的动机:前沿LRMs与人类游戏学习者的行为与大脑对齐
Botos Csaba, Sreejan Kumar, Austin Tudor David Andrews, Laurence Hunt, Chris Summerfield, Joshua B. Tenenbaum, Rui Ponte Costa, Marcelo G. Mattar, Momchil Tomov
机构
*
University of Oxford(牛津大学)
;
Columbia University(哥伦比亚大学)
;
New York University(纽约大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Harvard University(哈佛大学)