Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations
多买家市场中的战略谈判:基于可验证奖励的强化学习用于大语言模型谈判
Shuze Daniel Liu, Claire Chen, Jiabao Sean Xiao, Xin Chen, David Simchi-Levi
机构
*
Institute for Data, Systems, and Society, Massachusetts Institute of Technology(数据、系统与社会研究所,麻省理工学院)
;
Mitch Daniels School of Business, Purdue University(米奇·丹尼尔斯商学院,普渡大学)
;
The Division of Physics, Mathematics and Astronomy, California Institute of Technology(物理、数学与天文学部,加州理工学院)
;
Department of Computing and Mathematical Sciences, California Institute of Technology(计算与数学科学系,加州理工学院)
;
H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology(H. 米尔顿·斯图尔特工业与系统工程学院,佐治亚理工学院)
;
Department of Civil and Environmental Engineering, Operations Research Center, Massachusetts Institute of Technology(土木与环境工程系、运筹学中心,麻省理工学院)
WordVoice: Explicit and Decoupled Multi-Dimensional Word-Level Control for LLM-Based TTS
WordVoice:基于大语言模型的文本转语音中显式且解耦的多维词级控制
Sihang Nie, Jinxin Ji, Xiaofen Xing, Deyi Tuo, Chengbin Jin, Jialong Mai, Xiangmin Xu
机构
*
South China University of Technology(南方科技大学)
;
Huya Inc.(Huya公司)
;
Tongji University(同济大学)
;
The Hongkong polytechnic university(香港理工大学)
;
Foshan University(佛山大学)
Efficient Bethe-Salpeter Equation Calculations Based on Numerical Atomic Orbitals and Norm-Conserving Pseudopotentials: Dual-${\boldsymbol k}$-Mesh Strategy
Comments12 pages, 9 figures; Accepted for publication in Astronomy & Astrophysics. Code available at https://github.com/avikhagol/avica ; comments are welcome on the manuscript and code
Beyond Correctness: Enhancing Architectural Reasoning in Code LLMs via Scalable Labeling with Agentic Judgment
超越正确性:通过可扩展的智能体判断标注增强代码大模型的架构推理能力
Kirill Vasilevski, Ximing Dong, Benjamin Rombaut, Milad Soltany, Ruochen Deng, Jiahuei Lin, Arthur Leung, Dayi Lin, Boyuan Chen, Shaowei Wang, Ahmed E. Hassan
机构
*
Centre for Software Excellence, Huawei Canada(华为加拿大软件卓越中心)
;
Department of Computer Science, University of Manitoba, Canada(曼尼托巴大学计算机科学系)
;
School of Computing, Queen’s University, Canada(皇后大学计算科学学院)
Leveraging Metamemory Agent for Enhanced Data-Free Code Generation in Large Language Models
利用元记忆智能体增强大语言模型的无数据代码生成
Shengsheng Zhou, Shuai Wang, Liang Ding, Yibing Zhan, Yong Luo, Zheng He, Fu Lin, Dapeng Tao
机构
*
College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算机与数据科学学院,新加坡)
;
School of Computer Science, National Engineering Research Center for Multimedia Software, Wuhan University, Wuhan, China(多媒体软件国家工程研究中心计算机学院,武汉大学,中国)
;
The University of Sydney, Sydney, Australia(悉尼大学,澳大利亚)
;
Yunnan United Vision Technology Company Ltd., China(云南联合视界技术有限公司,中国)
;
School of Computer Science, Wuhan University, Wuhan, China(武汉大学计算机学院,中国)
;
School of Information Science and Engineering, Yunnan University, Kunming, China(云南大学信息科学与工程学院,中国)
An Experimental Design Approach to Evaluating Agentic AI's Autonomous Model Discovery
一种评估智能体人工智能自主模型发现的实验设计方法
Hao He, Xueying Liu, Chris J. Kuhlman, Xinwei Deng
机构
*
Department of Statistics, Virginia Tech(统计学系,弗吉尼亚理工学院)
;
Department of Statistical Science, Baylor University(统计科学系,贝勒大学)
;
Advanced Research Computing, Virginia Tech(高级研究计算,弗吉尼亚理工学院)