arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Princeton University(普林斯顿大学)

2026-04-16 至 2026-04-16 共收录 3
2510.19268 2026-04-16 cs.RO cs.LG

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

基于强化学习和上下文视觉-语言模型的分层DLO路由

Mingen Li, Houjian Yu, Yixuan Huang, Youngjin Hong, Hantao Ye, Changhyun Choi

机构 * Department of Electrical and Computer Engineering, University of Minnesota(明尼苏达大学电气与计算机工程系) Princeton University(普林斯顿大学)

AI总结 本文提出了一种自主分层框架,利用视觉-语言模型进行上下文高层推理,结合强化学习训练的低层技能,解决长视距DLO路由任务,实现92%的成功率。

Comments 8 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13305 2026-04-16 cs.CV

Bias at the End of the Score

评分末尾的偏见

Salma Abdel Magid, Grace Guo, Esin Tureci, Amaya Dharmasiri, Vikram V. Ramaswamy, Hanspeter Pfister, Olga Russakovsky

机构 * Princeton University(普林斯顿大学) Harvard University(哈佛大学)

AI总结 研究揭示奖励模型在文本到图像生成中因编码人口偏见导致性别和种族刻板印象强化及多样性崩溃的问题。

Comments Accepted to The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13192 2026-04-16 eess.SY cs.RO cs.SY

Synthesis and Deployment of Maximal Robust Control Barrier Functions through Adversarial Reinforcement Learning

通过对抗强化学习合成和部署最大鲁棒控制屏障函数

Donggeon David Oh, Duy P. Nguyen, Haimin Hu, Jaime Fernández Fisac

机构 * Department of Electrical and Computer Engineering, Princeton University(普林斯顿大学电气工程与计算机科学系) Department of Computer Science, Johns Hopkins University(约翰霍普金斯大学计算机科学系)

AI总结 本文提出了一种新的鲁棒控制屏障函数框架,用于一般非线性系统在有界不确定性下的安全控制,结合对抗强化学习实现鲁棒Q-CBF约束,验证了在倒立摆和四足机器人中的有效性。

Comments 8 pages, 2 figures. This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏