Sample Complexity of Average-Reward Q-Learning: From Single-agent to Federated Reinforcement Learning
平均奖励Q学习的样本复杂度:从单智能体到联邦强化学习
Yuchen Jiao, Jiin Woo, Gen Li, Gauri Joshi, Yuejie Chi
机构
*
CUHK Department of Statistics and Data Science, Chinese University of Hong Kong(中国香港中文大学统计与数据科学系)
;
CMU Department of Electrical and Computer Engineering, Carnegie Mellon University(卡内基梅隆大学电气与计算机工程系)
;
CUHK(中国香港中文大学)
;
CMU(卡内基梅隆大学)
;
Yale Department of Statistics and Data Science, Yale University(耶鲁大学统计与数据科学系)
Disagreement as Data: Reasoning Trace Analytics in Multi-Agent Systems
分歧作为数据:多智能体系统中的推理轨迹分析
Elham Tajik, Conrad Borchers, Bahar Shahrokhian, Sebastian Simon, Ali Keramati, Sonika Pal, Sreecharan Sankaranarayanan
机构
*
University at Albany(纽约州立大学阿尔巴尼分校)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Arizona State University(亚利桑那州立大学)
;
Le Mans University(勒曼大学)
;
University of California, Irvine(加州大学尔湾分校)
;
Indian Institute of Technology Bombay(印度班加罗尔理工学院)
;
Extuitive Inc. (Flagship Pioneering)(Extuitive公司(Flagship Pioneering))
机构
*
Carnegie Mellon University(卡内基梅隆大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Guangdong University of Technology(广东技术大学)
;
The University of Melbourne(墨尔本大学)
Let Me Try Again: Examining Replay Behavior by Tracing Students' Latent Problem-Solving Pathways
让我再试一次:通过追踪学生的潜在问题解决路径来研究重播行为
Shan Zhang, Siddhartha Pradhan, Ji-Eun Lee, Ashish Gurung, Anthony F. Botelho
机构
*
University of Florida(佛罗里达大学)
;
Worcester Polytechnic Institute(沃思黑普理工学院)
;
Singapore University of Technology and Design(新加坡科技与设计大学)
;
Carnegie Mellon University(卡内基梅隆大学)
机构
*
Princeton University(普林斯顿大学)
;
Princeton AI Lab(普林斯顿人工智能实验室)
;
Tsinghua University(清华大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Sydney(悉尼大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Pennsylvania State University(宾夕法尼亚州立大学)
;
University of Michigan(密歇根大学)
;
Oregon State University(俄勒冈州立大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Fudan University(复旦大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
The University of Hong Kong(香港大学)
;
University of California, Santa Barbara(加州大学圣芭芭拉分校)
;
University of California San Diego(加州大学圣地亚哥分校)
;
University of Edinburgh(爱丁堡大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
机构
*
Carnegie Mellon University(卡内基梅隆大学)
;
University of California, Berkeley(加州大学伯克利分校)
;
University of Texas, Austin(德克萨斯大学奥斯汀分校)
;
University of British Columbia(不列颠哥伦比亚大学)
Approximately Optimal Global Planning for Contact-Rich SE(2) Manipulation on a Graph of Reachable Sets
近似最优的接触丰富SE(2)操作在可达集图上的全局规划
Simin Liu, Tong Zhao, Bernhard Paus Graesdal, Peter Werner, Jiuguang Wang, John Dolan, Changliu Liu, Tao Pang
机构
*
Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人学院)
;
Robotics and AI Institute(机器人与人工智能研究所)
;
CSAIL, Massachusetts Institute of Technology(麻省理工学院计算机科学与人工智能实验室)