arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4116 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4116 篇

2508.05619 2025-08-08 cs.AI nlin.AO physics.bio-ph physics.comp-ph physics.hist-ph 57%

The Missing Reward: Active Inference in the Era of Experience

Bo Wen

机构 * IBM T.J. Watson Research Center(IBM TJ沃森研究中心)

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03910 2025-08-07 q-fin.CP cs.LG 57%

Comparing Normalization Methods for Portfolio Optimization with Reinforcement Learning

Caio de Souza Barbosa Costa, Anna Helena Reali Costa

机构 * Escola Politécnica Universidade de São Paulo (USP)(圣保罗大学理工学院)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03012 2025-08-07 cs.SE cs.AI 57%

Tool-integrated Reinforcement Learning for Repo Deep Search

Zexiong Ma, Chao Peng, Qunhong Zeng, Pengfei Gao, Yanzhen Zou, Bing Xie

机构 * Peking University(北京大学) ByteDance(字节跳动) Beijing Institute of Technology(北京理工大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05858 2025-08-07 cs.LG cs.SY eess.SY 57%

Distributional Soft Actor-Critic with Three Refinements

Jingliang Duan, Wenxuan Wang, Liming Xiao, Jiaxin Gao, Shengbo Eben Li, Chang Liu, Ya-Qin Zhang, Bo Cheng, Keqiang Li

机构 * School of Mechanical Engineering, University of Science and Technology Beijing(北京科技大学机械工程学院) School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动性学院) College of Engineering, Peking University(北京大学工程学院) Institute for AI Industry Research, Tsinghua University(清华大学人工智能产业研究院)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

Comments Title updated in this version. The previous version was titled "DSAC-T: Distributional Soft Actor-Critic With Three Refinements". Added a footnote with the BibTeX entry for the published journal version. No other major changes

Journal ref IEEE Trans. PAMI,47(5): 3935-3946,2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03129 2025-08-06 cs.RO 57%

Safety-Aware Imitation Learning via MPC-Guided Disturbance Injection

Le Qiu, Yusuf Umut Ciftci, Somil Bansal

机构 * Department of Electrical Engineering, Tsinghua University(电子工程系,清华大学) Department of Electrical and Computer Engineering, University of Southern California(电气与计算机工程系,南加州大学) Department of Aeronautics and Astronautics, Stanford University(航空与航天工程系,斯坦福大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14476 2025-08-06 cs.AI 57%

Learning telic-controllable state representations

Nadav Amir, Stas Tiomkin

机构 * Princeton Neuroscience Institute, Princeton University(普林斯顿神经科学研究所,普林斯顿大学) Department of Computer Science, Texas Tech University(计算机科学系,德克萨斯技术大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI

Comments Published in Proceedings of the 47th Annual Meeting of the Cognitive Science Society

Journal ref Proceedings of the Annual Meeting of the Cognitive Science Society, 47 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04974 2025-08-04 cs.CV 57%

ReAlign: Bilingual Text-to-Motion Generation via Step-Aware Reward-Guided Alignment

Wanjiang Weng, Xiaofeng Tan, Hongsong Wang, Pan Zhou

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.CV

Comments We believe that there are some areas in the manuscript that require further improvement, and out of our commitment to refining this work, we have decided to withdraw our manuscript after careful deliberation and discussion

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23562 2025-08-01 cs.LG cs.AR 57%

Hardware-Aware Fine-Tuning of Spiking Q-Networks on the SpiNNaker2 Neuromorphic Platform

Sirine Arfa, Bernhard Vogginger, Christian Mayr

机构 * Centre for Tactile Internet with Human-in-the-Loop (CeTI)(触觉互联网与人环路中心)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

Comments 8 pages, 5 figures, 3 tables

Journal ref ACM ICONS 2025 - International Conference on Neuromorphic Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06684 2025-08-01 cs.RO cs.MA 57%

SDHN: Skewness-Driven Hypergraph Networks for Enhanced Localized Multi-Robot Coordination

Delin Zhao, Yanbo Shan, Chang Liu, Shenghang Lin, Yingxin Shou, Bin Xu

机构 * School of Automation, Northwestern Polytechnical University, Xi'an, 710072, China(西北工业大学自动化学院) School of Automation, Southeast University, Nanjing, 210096, China(东南大学自动化学院)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09722 2025-07-29 cs.LG cs.SY eess.SY stat.ML 57%

The Pitfalls of Imitation Learning when Actions are Continuous

Max Simchowitz, Daniel Pfrommer, Ali Jadbabaie

机构 * CMU(卡内基梅隆大学) MIT(麻省理工学院)

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.LG

Comments 98 pages, 2 figures, updated proof sketch

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18113 2025-07-25 cs.LG 57%

Policy Disruption in Reinforcement Learning:Adversarial Attack with Large Language Models and Critical State Identification

Junyong Jiang, Buwei Tian, Chenxing Xu, Songze Li, Lu Dong

机构 * School of Cyber Science and Engineering(网络科学与工程学院)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16941 2025-07-24 cs.RO 57%

Multi-agent Reinforcement Learning for Robotized Coral Reef Sample Collection

Daniel Correa, Tero Kaarlela, Jose Fuentes, Paulo Padrao, Alain Duran, Leonardo Bobadilla

机构 * Knight Foundation School of Computing and Information Sciences, Florida International University(凯尼恩基金会计算与信息科学学院,佛罗里达国际大学) College of Arts, Sciences & Education, Biological Sciences, Florida International University(艺术、科学与教育学院,生物科学,佛罗里达国际大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15073 2025-07-22 cs.LG 57%

Reinforcement Learning for Flow-Matching Policies

Samuel Pfrommer, Yixiao Huang, Somayeh Sojoudi

机构 * Department of Electrical Engineering and Computer Sciences University of California, Berkeley(电气工程与计算机科学系 加州大学伯克利分校)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01349 2025-07-17 cs.AI cs.NE 57%

Life, uh, Finds a Way: Hyperadaptability by Behavioral Search

Alex Baranski, Jun Tani

机构 * Okinawa Institute of Science and Technology(冲绳科学技术大学院)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI

Comments 39 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06151 2025-07-16 cs.CV 57%

Biomechanics-Guided Residual Approach to Generalizable Human Motion Generation and Estimation

Zixi Kang, Xinghan Wang, Yadong Mu

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11173 2025-07-16 cs.LG 57%

Real-Time Bayesian Detection of Drift-Evasive GNSS Spoofing in Reinforcement Learning Based UAV Deconfliction

Deepak Kumar Panda, Weisi Guo

机构 * Centre for Connected and Assured Autonomy, Faculty of Engineering and Applied Sciences, Cranfield University(连接与可信自主性中心,工程与应用科学学院,克兰菲尔德大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10284 2025-07-15 cs.RO cs.MA 57%

Prompt Informed Reinforcement Learning for Visual Coverage Path Planning

Venkat Margapuri

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10030 2025-07-15 cs.RO 57%

Finetuning Deep Reinforcement Learning Policies with Evolutionary Strategies for Control of Underactuated Robots

Marco Calì, Alberto Sinigaglia, Niccolò Turcato, Ruggero Carli, Gian Antonio Susto

机构 * Department of Information Engineering, University of Padova(信息工程系,帕多瓦大学) Human-Inspired Technology Research Center, University of Padova(人机协同技术研究中心,帕多瓦大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08112 2025-07-14 cs.RO 57%

Imitation Learning for Obstacle Avoidance Using End-to-End CNN-Based Sensor Fusion

Lamiaa H. Zain, Hossam H. Ammar, Raafat E. Shalaby

机构 * School of Engineering(工程学院) Applied Science Nile University(应用科学尼罗大学) Computer Science University of Hertfordshire(计算机科学赫尔德福德郡大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05720 2025-07-09 cs.LG cs.CL 57%

MobileGUI-RL: Advancing Mobile GUI Agent through Reinforcement Learning in Online Environment

Yucheng Shi, Wenhao Yu, Zaitang Li, Yonglin Wang, Hongming Zhang, Ninghao Liu, Haitao Mi, Dong Yu

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments 17 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10251 2025-07-09 cs.RO 57%

SRT-H: A Hierarchical Framework for Autonomous Surgery via Language Conditioned Imitation Learning

Ji Woong Kim, Juo-Tung Chen, Pascal Hansen, Lucy X. Shi, Antony Goldenberg, Samuel Schmidgall, Paul Maria Scheikl, Anton Deguet, Brandon M. White, De Ru Tsai, Richard Cha, Jeffrey Jopling, Chelsea Finn, Axel Krieger

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06495 2025-07-08 cs.LG 57%

Boosting Robustness in Preference-Based Reinforcement Learning with Dynamic Sparsity

Calarina Muslimani, Bram Grooten, Deepak Ranganatha Sastry Mamillapalli, Mykola Pechenizkiy, Decebal Constantin Mocanu, Matthew E. Taylor

机构 * University of Alberta(阿尔伯塔大学) Eindhoven University of Technology(埃因霍温理工大学) University of Luxembourg(卢森堡大学) Alberta Machine Intelligence Institute(阿尔伯塔人工智能研究所)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04164 2025-07-04 cs.LG 57%

MInCo: Mitigating Information Conflicts in Distracted Visual Model-based Reinforcement Learning

Shiguang Sun, Hanbo Zhang, Zeyang Liu, Xinrui Yang, Lipeng Wan, Xingyu Chen, Xuguang Lan

机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(人机混合增强智能国家重点实验室) National Engineering Research Center for Visual Information and Applications(视觉信息与应用国家工程研究中心) Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院) Xi’an Jiaotong University(西安交通大学) National University of Singapore(新加坡国立大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21899 2025-06-30 cs.LG 57%

Advancements and Challenges in Continual Reinforcement Learning: A Comprehensive Review

Amara Zuffer, Michael Burke, Mehrtash Harandi

机构 * Electrical and Computer Science Engineering Department, Faculty of Engineering, Monash University, Clayton, Victoria, Australia(莫纳什大学工程学院电气与计算机科学工程系)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments 65 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18126 2025-06-24 cs.AI 57%

Decentralized Consensus Inference-based Hierarchical Reinforcement Learning for Multi-Constrained UAV Pursuit-Evasion Game

Xiang Yuming, Li Sizhao, Li Rongpeng, Zhao Zhifeng, Zhang Honggang

机构 * College of Information Science and Electronic Engineering, Zhejiang University(浙江大学信息科学与电子工程学院) Zhejiang Lab(浙江省实验室) City University of Macau(澳门城市大学)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10048 2025-06-24 cs.SD cs.AI eess.AS 57%

Audio-Driven Reinforcement Learning for Head-Orientation in Naturalistic Environments

Wessel Ledder, Yuzhen Qin, Kiki van der Heijden

机构 * Radboud University(拉德堡德大学) Donders Institute(多纳德研究所)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI

Comments Accepted at ICASSP 2025

Journal ref Proc. IEEE Int. Conf. Acoust., Speech, Signal Process. (ICASSP), pp. 1-5, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16336 2025-06-23 cs.RO cs.MA 57%

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections

Yiou Huang

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15132 2025-06-19 cs.RO 57%

Booster Gym: An End-to-End Reinforcement Learning Framework for Humanoid Robot Locomotion

Yushi Wang, Penghui Chen, Xinyu Han, Feng Wu, Mingguo Zhao

机构 * Department of Automation, Tsinghua University(自动化系,清华大学) Booster Robotics Technology Co., Ltd(Booster机器人技术有限公司)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14973 2025-06-18 cs.AI 57%

Behaviour Discovery and Attribution for Explainable Reinforcement Learning

Rishav Rishav, Somjit Nath, Vincent Michalski, Samira Ebrahimi Kahou

机构 * University of Calgary(卡尔加里大学) Mila McGill University(麦吉尔大学) Université de Montréal(蒙特利尔大学) CIFAR AI Chair(CIFAR AI 职位)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12161 2025-06-17 cs.LG stat.ML 57%

Meta-Learning and Synthetic Data for Automated Pretraining and Finetuning

Fabio Ferreira

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.LG

Comments PhD thesis

详情

展开后加载摘要…

URL PDF HTML 收藏