MATH-PT: A Math Reasoning Benchmark for European and Brazilian Portuguese
MATH-PT:一个针对欧洲葡萄牙语和巴西葡萄牙语的数学推理基准数据集
Tiago Teixeira, Ana Carolina Erthal, Juan Belieni, Beatriz Canaverde, Diego Mesquita, Miguel Faria, Eliezer de Souza da Silva, André F. T. Martins
机构
*
Instituto Superior Técnico, Universidade de Lisboa(里斯本大学理工学院)
;
Fundação Getulio Vargas(古特曼基金会)
;
Instituto de Telecomunicações(电信研究所)
;
Universidade de Coimbra, CISUC/LASI, DEI(科英布拉大学,CISUC/LASI,DEI)
;
Basque Center for Applied Mathematics(巴斯克应用数学中心)
机构
*
Shenyang Institute of Computing Technology, Chinese Academy of Sciences(中国科学院沈阳计算技术研究所)
;
University of Chinese Academy of Sciences 3 ByteDance 4 Westlake University(中国科学院大学 3 字节跳动 4 西湖大学)
;
Key Laboratory of Computing Power Network(计算功率网络重点实验室)
;
Information Security, Ministry of Education, Shandong Computer Science Center (National Supercomputer Center in Jinan), Qilu University of Technology (Shandong Academy of Sciences)(信息安全,教育部,山东计算机科学中心(济南国家超级计算中心),齐鲁工业大学(山东省科学院))
Rethinking Entropy Interventions in RLVR: An Entropy Change Perspective
重新思考RLVR中的熵干预:从熵变化视角
Zhezheng Hao, Hong Wang, Haoyang Liu, Jian Luo, Jiarui Yu, Hande Dong, Qiang Lin, Can Wang, Jiawei Chen
机构
*
State Key Laboratory of Blockchain and Data Security(区块链与数据安全国家重点实验室)
;
Zhejiang University(浙江大学)
;
Tencent(腾讯)
;
Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and Data Security(杭州高新技术区(滨江)区块链与数据安全研究院)
CommentsSubstantially revised version consolidating the paper as a formal SAT-Graph API specification: clarifies Probability Isolation and post-anchoring determinism, broadens semantic anchoring to open and thematic legal queries, refines the data models and temporal primitives, and strengthens the use cases, limitations, and bibliography
How Do AI Agents Spend Your Money? Analyzing and Predicting Token Consumption in Agentic Coding Tasks
AI代理如何花费你的钱?分析和预测代理编码任务中的令牌消耗
Longju Bai, Zhemin Huang, Xingyao Wang, Jiao Sun, Rada Mihalcea, Erik Brynjolfsson, Alex Pentland, Jiaxin Pei
机构
*
University of Michigan(密歇根大学)
;
Stanford University(斯坦福大学)
;
All Hands AI
;
Google Deepmind(谷歌DeepMind)
;
Microsoft AI(微软AI)
;
Massachusetts Institute of Technology(麻省理工学院)
CommentsThis is an extended version of the paper with the same title that will appear in KR 2026, and which contains a technical appendix with proof details
Retrieval-Augmented Multimodal Model for Fake News Detection
增强检索的多模态模型用于虚假新闻检测
Yiheng Li, Weihai Lu, Hanyi Yu, Yue Wang
机构
*
University of International Business and Economics(国际商务经济大学)
;
Peking University(北京大学)
;
University of Southern California(南加州大学)
;
Upstart Holdings, Inc.(Upstart Holdings公司)
Adaptive and Fine-grained Module-wise Expert Pruning for Efficient LoRA-MoE Fine-Tuning
自适应和细粒度模块级专家剪枝用于高效的LoRA-MoE微调
Weihang Li, Jianchun Liu, Hongli Xu
机构
*
School of Computer Science and Technology, University of Science and Technology of China, China(计算机科学与技术学院,中国科学技术大学)
;
Suzhou Institute for Advanced Research, University of Science and Technology of China, China(苏州先进研究院,中国科学技术大学)
Virtual-reality based patient-specific simulation of spine surgical procedures: A fast, highly automated and high-fidelity system for surgical education and planning
基于虚拟现实的患者特异性脊柱手术模拟:一种快速、高度自动化和高保真的系统,用于手术教育和规划
Raj Kumar Ranabhat, Tayler D Ross, Tony Jiao, Jeremie Larouche, Joel Finkelstein, Michael Hardisty
机构
*
Holland Bone and Joint Program, Sunnybrook Research Institute(霍尔德骨科与关节程序,圣布鲁诺研究学院)
;
Division of Spine Surgery, Sunnybrook Health Sciences Centre(脊柱外科部,圣布鲁诺健康科学中心)
;
Division of Orthopaedic Surgery, Department of Surgery, University of Toronto(骨科部,外科部,多伦多大学)
CommentsFull version of the work accepted as a short paper at the 34th ACM Conference on User Modeling, Adaptation and Personalization (UMAP '26). 9 pages, 4 figures, 5 tables
Journal refProceedings of the 34th ACM Conference on User Modeling, Adaptation and Personalization (UMAP '26), June 08--11, 2026, Gothenburg, Sweden
A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring
基于决策理论的隐写术形式化及其在大语言模型监控中的应用
Usman Anwar, Julianna Piskorz, David D. Baek, David Africa, Jim Weatherall, Max Tegmark, Christian Schroeder de Witt, Mihaela van der Schaar, David Krueger
机构
*
University of Cambridge(剑桥大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
UK AI Safety Institute(英国人工智能安全研究所)
;
University of Oxford(牛津大学)
;
Mila, University of Montreal(蒙特利尔大学米尔人工智能实验室)