arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

共收录 7719 信号源:cs.CV, cs.GR, cs.RO

1. 点云 7719 篇

2501.14502 2025-01-27 cs.RO cs.CV 62%

LiDAR-Based Vehicle Detection and Tracking for Autonomous Racing

Marcello Cellina, Matteo Corno, Sergio Matteo Savaresi

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08593 2025-01-16 cs.RO cs.CV 62%

Image-to-Force Estimation for Soft Tissue Interaction in Robotic-Assisted Surgery Using Structured Light

Jiayin Wang, Mingfeng Yao, Yanran Wei, Xiaoyu Guo, Ayong Zheng, Weidong Zhao

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.15975 2025-01-15 cs.RO cs.CV 62%

Analyzing Infrastructure LiDAR Placement with Realistic LiDAR Simulation Library

Xinyu Cai, Wentao Jiang, Runsheng Xu, Wenquan Zhao, Jiaqi Ma, Si Liu, Yikang Li

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments 7 pages, 6 figures, accepted to the IEEE International Conference on Robotics and Automation (ICRA'23)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05472 2025-01-13 cs.CV cs.LG cs.RO 62%

The 2nd Place Solution from the 3D Semantic Segmentation Track in the 2024 Waymo Open Dataset Challenge

Qing Wu

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10748 2025-01-07 cs.CV cs.GR cs.LG physics.flu-dyn 62%

A Pioneering Neural Network Method for Efficient and Robust Fluid Simulation

Yu Chen, Shuai Zheng, Nianyi Wang, Menglong Jin, Yan Chang

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.GR

Comments This paper has been accepted by AAAI Conference on Artificial Intelligence (AAAI-25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17226 2024-12-24 cs.CV cs.RO 62%

OLiDM: Object-aware LiDAR Diffusion Models for Autonomous Driving

Tianyi Yan, Junbo Yin, Xianpeng Lang, Ruigang Yang, Cheng-Zhong Xu, Jianbing Shen

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments AAAI 2025, https://yanty123.github.io/OLiDM

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13662 2024-12-19 cs.CV cs.AI cs.LG cs.RO 62%

When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?

Tongzhou Mu, Zhaoyang Li, Stanisław Wiktor Strzelecki, Xiu Yuan, Yunchao Yao, Litian Liang, Hao Su

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted by The 39th Annual AAAI Conference on Artificial Intelligence (AAAI 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11489 2024-12-17 cs.CV cs.AI cs.LG cs.RO 62%

HGSFusion: Radar-Camera Fusion with Hybrid Generation and Synchronization for 3D Object Detection

Zijian Gu, Jianwei Ma, Yan Huang, Honghao Wei, Zhanye Chen, Hui Zhang, Wei Hong

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments 12 pages, 8 figures, 7 tables. Accepted by AAAI 2025 , the 39th Annual AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05426 2024-12-10 cs.RO cs.AI cs.CV cs.LG 62%

What's the Move? Hybrid Imitation Learning via Salient Points

Priya Sundaresan, Hengyuan Hu, Quan Vuong, Jeannette Bohg, Dorsa Sadigh

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15682 2024-12-06 cs.CV cs.RO 62%

RANSAC Back to SOTA: A Two-stage Consensus Filtering for Real-time 3D Registration

Pengcheng Shi, Shaocheng Yan, Yilin Xiao, Xinyi Liu, Yongjun Zhang, Jiayuan Li

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments 8 pages, 9 figures

Journal ref IEEE Robotics and Automation Letters, vol.9, no.12, pp.11881-11888, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01539 2024-12-03 cs.CV cs.RO 62%

The Bare Necessities: Designing Simple, Effective Open-Vocabulary Scene Graphs

Christina Kassab, Matías Mattamala, Sacha Morin, Martin Büchner, Abhinav Valada, Liam Paull, Maurice Fallon

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01299 2024-12-03 cs.CV cs.RO 62%

Cross-Modal Visual Relocalization in Prior LiDAR Maps Utilizing Intensity Textures

Qiyuan Shen, Hengwang Zhao, Weihao Yan, Chunxiang Wang, Tong Qin, Ming Yang

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.06963 2024-12-03 cs.GR cs.AI cs.CV cs.LG 62%

ELMO: Enhanced Real-time LiDAR Motion Capture through Upsampling

Deok-Kyeong Jang, Dongseok Yang, Deok-Yun Jang, Byeoli Choi, Donghoon Shin, Sung-hee Lee

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.GR

Comments published at ACM Transactions on Graphics (Proc. SIGGRAPH ASIA), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18476 2024-11-28 cs.RO cs.CV 62%

A comparison of extended object tracking with multi-modal sensors in indoor environment

Jiangtao Shuai, Martin Baerveldt, Manh Nguyen-Duc, Anh Le-Tuan, Manfred Hauswirth, Danh Le-Phuoc

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08850 2024-11-20 cs.CV cs.RO 62%

LiDAR-BEVMTN: Real-Time LiDAR Bird's-Eye View Multi-Task Perception Network for Autonomous Driving

Sambit Mohapatra, Senthil Yogamani, Varun Ravi Kumar, Stefan Milz, Heinrich Gotzig, Patrick Mäder

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted for publication at IEEE Transactions on Intelligent Transportation Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.02389 2024-11-19 cs.CV cs.AI cs.RO 62%

Multi-modal Situated Reasoning in 3D Scenes

Xiongkun Linghu, Jiangyong Huang, Xuesong Niu, Xiaojian Ma, Baoxiong Jia, Siyuan Huang

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted by NeurIPS 2024 Datasets and Benchmarks Track. Project page: https://msr3d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10203 2024-11-18 cs.CV cs.RO 62%

Learning Generalizable 3D Manipulation With 10 Demonstrations

Yu Ren, Yang Cong, Ronghan Chen, Jiahao Long

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.15364 2024-11-13 cs.RO cs.CV 62%

WildScenes: A Benchmark for 2D and 3D Semantic Segmentation in Large-scale Natural Environments

Kavisha Vidanapathirana, Joshua Knights, Stephen Hausler, Mark Cox, Milad Ramezani, Jason Jooste, Ethan Griffiths, Shaheer Mohamed, Sridha Sridharan, Clinton Fookes, Peyman Moghadam

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted in the The International Journal of Robotics Research (IJRR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05311 2024-11-11 cs.CV cs.RO 62%

ZOPP: A Framework of Zero-shot Offboard Panoptic Perception for Autonomous Driving

Tao Ma, Hongbin Zhou, Qiusheng Huang, Xuemeng Yang, Jianfei Guo, Bo Zhang, Min Dou, Yu Qiao, Botian Shi, Hongsheng Li

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted by NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05003 2024-11-08 cs.CV cs.AI cs.GR cs.LG 62%

ReCapture: Generative Video Camera Controls for User-Provided Videos using Masked Video Fine-Tuning

David Junhao Zhang, Roni Paiss, Shiran Zada, Nikhil Karnad, David E. Jacobs, Yael Pritch, Inbar Mosseri, Mike Zheng Shou, Neal Wadhwa, Nataniel Ruiz

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.GR

Comments project page: https://generative-video-camera-controls.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.10159 2024-11-05 cs.CV cs.LG cs.RO 62%

RAPiD-Seg: Range-Aware Pointwise Distance Distribution Networks for 3D LiDAR Segmentation

Li Li, Hubert P. H. Shum, Toby P. Breckon

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments ECCV 2024 (Oral); 18 pages, 6 figures, 7 tables; Code at https://github.com/l1997i/rapid_seg

Journal ref Eur. Conf. Comput. Vis. (ECCV 2024 ORAL)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00600 2024-11-04 cs.CV cs.AI cs.RO 62%

On Deep Learning for Geometric and Semantic Scene Understanding Using On-Vehicle 3D LiDAR

Li Li

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments PhD thesis (Durham University, Computer Science), 149 pages (the 2024 BMVA Sullivan Doctoral Thesis Prize runner-up). Includes published content from arXiv:2407.10159 (ECCV 2024 ORAL), arXiv:2303.11203 (CVPR 2023), and arXiv:2406.10068 (3DV 2021), with minor revisions to the examined version: https://etheses.dur.ac.uk/15738/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22707 2024-10-31 cs.RO cs.AI cs.CV 62%

Robotic State Recognition with Image-to-Text Retrieval Task of Pre-Trained Vision-Language Model and Black-Box Optimization

Kento Kawaharazuka, Yoshiki Obinata, Naoaki Kanazawa, Kei Okada, Masayuki Inaba

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted at Humanoids2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15250 2024-10-24 cs.CV cs.RO 62%

D2S: Representing sparse descriptors and 3D coordinates for camera relocalization

Bach-Thuan Bui, Huy-Hoang Bui, Dinh-Tuan Tran, Joo-Ho Lee

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted to IEEE Robotics and Automation Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17633 2024-10-22 cs.CV cs.AI cs.RO 62%

UADA3D: Unsupervised Adversarial Domain Adaptation for 3D Object Detection with Sparse LiDAR and Large Domain Gaps

Maciej K Wozniak, Mattias Hansson, Marko Thiel, Patric Jensfelt

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted for IEEE RA-L 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13860 2024-10-18 cs.CV cs.RO 62%

VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Runsen Xu, Zhiwei Huang, Tai Wang, Yilun Chen, Jiangmiao Pang, Dahua Lin

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments CoRL 2024 Camera Ready. 25 pages. A novel zero-shot 3D visual grounding framework based solely on 2D images

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12995 2024-10-18 cs.RO cs.CV 62%

Configurable Embodied Data Generation for Class-Agnostic RGB-D Video Segmentation

Anthony Opipari, Aravindhan K Krishnan, Shreekant Gayaka, Min Sun, Cheng-Hao Kuo, Arnie Sen, Odest Chadwicke Jenkins

专题命中 点云 :3D reconstruction(abstract);分类 cs.CV、cs.RO

Comments Accepted in IEEE Robotics and Automation Letters October 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.10115 2024-10-16 cs.CV cs.LG cs.RO 62%

Shelf-Supervised Cross-Modal Pre-Training for 3D Object Detection

Mehar Khurana, Neehar Peri, James Hays, Deva Ramanan

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments The first two authors contributed equally. This work has been accepted to the Conference on Robot Learning (CoRL) 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18722 2024-10-15 cs.RO cs.CV 62%

Towards Open-World Grasping with Large Vision-Language Models

Georgios Tziafas, Hamidreza Kasaei

专题命中 点云 :spatial understanding(abstract);分类 cs.CV、cs.RO

Comments 8th Conference on Robot Learning (CoRL 2024), Munich, Germany

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.08365 2024-10-14 cs.RO cs.CV 62%

Are We Ready for Real-Time LiDAR Semantic Segmentation in Autonomous Driving?

Samir Abou Haidar, Alexandre Chariot, Mehdi Darouich, Cyril Joly, Jean-Emmanuel Deschaud

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

Comments Accepted to IROS 2024 PPNIV Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏