Hierarchical Reinforcement Learning for Sparse-Reward Search in Commutative Algebra
用于交换代数中稀疏奖励搜索的分层强化学习
Giorgi Butbaia, Paul Orland, Coco Huang, Davide Passaro, Lucas Fagan, Michele Tarquini, Hailong Dao, David Eisenbud, Ali Shehper, Sergei Gukov
机构
*
California Institute of Technology(加州理工学院)
;
Temple University(天普大学)
;
University of Kansas(堪萨斯大学)
;
University of California, Berkeley(加州大学伯克利分校)
The Two-Hump Problem: Bridging the Difficulty Gap in Mathematical Reinforcement Learning
双峰问题:弥合数学强化学习中的难度差距
Lucas Fagan, Michele Tarquini, Ali Shehper, Maksymilian Manko, Angus Gruen, Coco Huang, Giorgi Butbaia, Davide Passaro, Sergei Gukov
机构
*
Department of Mathematics, California Institute of Technology(加州理工学院数学系)
;
Institute of Mathematics, University of Zurich(苏黎世大学数学研究所)
;
Zero Latency Labs(零延迟实验室)
;
Department of Mathematics, Temple University(天普大学数学系)
CommentsAccepted to the 2026 IEEE International Conference on Robotics and Automation (ICRA 2026). Copyright transferred to IEEE. Sample code for the navigation example with CBF-RL reward core construction can be found at https://github.com/lzyang2000/cbf-rl-navigation-demo
Flow Matching for Efficient and Scalable Data Assimilation
用于高效可扩展数据同化的流匹配
Taos Transue, Bohan Chen, So Takao, Bao Wang
机构
*
The Computing and Mathematical Sciences Department, California Institute of Technology(加州理工学院计算与数学科学系)
;
Department of Mathematics and Scientific Computing and Imaging Institute, University of Utah(犹他大学数学与科学计算系和成像研究所)
Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems
基于高斯过程策略迭代与加性Schwarz加速的正向和逆向HJB及平均场博弈问题
Xianjin Yang, Jingguo Zhang
机构
*
Department of Computing and Mathematical Sciences, California Institute of Technology, CA, USA(计算与数学科学系,加州理工学院,CA,美国)
;
Department of Mathematics and Risk Management Institute, National University of Singapore, Singapore(数学与风险管理研究所,新加坡国立大学,新加坡)
SHIELD: Safety on Humanoids via CBFs In Expectation on Learned Dynamics
SHIELD: 基于学习动力学期望的控制障碍函数实现人形机器人安全
Lizhi Yang, Blake Werner, Ryan K. Cosner, David Fridovich-Keil, Preston Culbertson, Aaron D. Ames
机构
*
Mechanical and Civil Engineering, California Institute of Technology(加州理工学院机械与土木工程系)
;
Aerospace Engineering and Engineering Mechanics, UT Austin(德克萨斯大学奥斯汀分校航空航天工程与工程力学系)
;
Computer Science, Cornell University(康奈尔大学计算机科学系)
CommentsAccepted to the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025). Copyright transferred to IEEE. Video at https://youtu.be/-Qv1wR4jfj4