arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

共收录 19012 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 4394 篇

2407.13764 2025-10-17 cs.CV 57%

Shape of Motion: 4D Reconstruction from a Single Video

Qianqian Wang, Vickie Ye, Hang Gao, Weijia Zeng, Jake Austin, Zhengqi Li, Angjoo Kanazawa

机构 * UC Berkeley(加州大学伯克利分校) Google DeepMind(谷歌DeepMind) UC San Diego(加州大学圣地亚哥分校) Adobe Research(Adobe研究)

专题命中 三维重建 :novel view synthesis(abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14234 2025-10-17 cs.RO cs.SY eess.SY 57%

Prescribed Performance Control of Deformable Object Manipulation in Spatial Latent Space

Ning Han, Gu Gong, Bin Zhang, Yuexuan Xu, Bohan Yang, Yunhui Liu, David Navarro-Alarcon

机构 * Department of Mechanical Engineering, The Hong Kong Polytechnic University(机械工程系,香港理工大学) T Stone Robotics Institute, Department of Mechanical and Automation Eng., The Chinese University of Hong Kong(T Stone机器人研究所,机械与自动化工程系,香港中文大学)

专题命中 三维重建 :point cloud(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07944 2025-10-17 cs.CV 57%

CVD-STORM: Cross-View Video Diffusion with Spatial-Temporal Reconstruction Model for Autonomous Driving

Tianrui Zhang, Yichen Liu, Zilin Guo, Yuxin Guo, Jingcheng Ni, Chenjing Ding, Dan Xu, Lewei Lu, Zehuan Wu

机构 * Sensetime Research(商汤科技研究院) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 三维重建 :Gaussian Splatting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13678 2025-10-16 cs.CV 57%

FlashWorld: High-quality 3D Scene Generation within Seconds

Xinyang Li, Tengfei Wang, Zixiao Gu, Shengchuan Zhang, Chunchao Guo, Liujuan Cao

机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学) Tencent(腾讯) Yes Lab, Fudan University(复旦大学Yes实验室)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments Project Page: https://imlixinyang.github.io/FlashWorld-Project-Page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13652 2025-10-16 cs.CV 57%

EditCast3D: Single-Frame-Guided 3D Editing with Video Propagation and View Selection

Huaizhi Qu, Ruichen Zhang, Shuqing Luo, Luchao Qi, Zhihao Zhang, Xiaoming Liu, Roni Sengupta, Tianlong Chen

机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) Michigan State University(密歇根州立大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05483 2025-10-15 cs.CV 57%

Veriserum: A dual-plane fluoroscopic dataset with knee implant phantoms for deep learning in medical imaging

Jinhao Wang, Florian Vogl, Pascal Schütz, Saša Ćuković, William R. Taylor

机构 * ETH Zürich, Laboratory for Movement Biomechanics, Institute for Biomechanics, Switzerland(苏黎世联邦理工学院,运动生物力学实验室,生物力学研究所,瑞士)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments This work has been accepted at MICCAI 2025

Journal ref In: Medical Image Computing and Computer-Assisted Intervention (MICCAI 2025), Lecture Notes in Computer Science (LNCS), Springer, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05864 2025-10-15 cs.CV 57%

CryoFastAR: Fast Cryo-EM Ab Initio Reconstruction Made Easy

Jiakai Zhang, Shouchen Zhou, Haizhao Dai, Xinhang Liu, Peihao Wang, Zhiwen Fan, Yuan Pei, Jingyi Yu

机构 * ShanghaiTech University(上海科技大學) Cellverse, Co., Ltd(Cellverse公司) HKUST(香港科技大学) UT Austin(德克萨斯大学奥斯汀分校)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11687 2025-10-14 cs.CV 57%

Beyond 'Templates': Category-Agnostic Object Pose, Size, and Shape Estimation from a Single View

Jinyu Zhang, Haitao Lin, Jiashu Hou, Xiangyang Xue, Yanwei Fu

机构 * Fudan University(复旦大学) Shanghai Innovation Institute(上海创新研究院)

专题命中 三维重建 :point cloud(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11632 2025-10-14 cs.CV cs.AI cs.LG 57%

NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection

Krittin Chaowakarn, Paramin Sangwongngam, Nang Htet Htet Aung, Chalie Charoenlarpnopparut

机构 * The School of Information, Computer, and Communication Technology, Sirindhorn International Institute of Technology, Thammasat University(信息、计算机与通信技术学院,Sirindhorn国际技术学院,泰国朱拉隆梭大学) National Electronics and Computer Technology Center, National Science and Technology Development Agency(国家电子与计算机技术中心,国家科学技术发展局) Department of Electrical Engineering, Faculty of Engineering, Chulalongkorn University(电气工程系,工程学院,朱拉隆梭大学)

专题命中 三维重建 :point cloud(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07723 2025-10-14 cs.CV 57%

SyncHuman: Synchronizing 2D and 3D Generative Models for Single-view Human Reconstruction

Wenyue Chen, Peng Li, Wangguandong Zheng, Chengfeng Zhao, Mengfei Li, Yaolong Zhu, Zhiyang Dou, Ronggang Wang, Yuan Liu

机构 * PKU(北京大学) HKUST(香港科技大学) SEU(上海师范大学) HKU(香港大学)

专题命中 三维重建 :3D generation(abstract);分类 cs.CV

Comments NeurIPS 2025 https://xishuxishu.github.io/SyncHuman.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10406 2025-10-14 cs.CV cs.AI cs.LG 57%

Mesh-Gait: A Unified Framework for Gait Recognition Through Multi-Modal Representation Learning from 2D Silhouettes

Zhao-Yang Wang, Jieneng Chen, Jiang Liu, Yuxiang Guo, Rama Chellappa

机构 * Johns Hopkins University(约翰霍普金斯大学) Advanced Micro Devices, Inc.(先进微器件公司)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01770 2025-10-14 cs.RO cs.AI cs.LG 57%

Robot Learning with Super-Linear Scaling

Marcel Torne, Arhan Jain, Jiayi Yuan, Vidaaranya Macha, Lars Ankile, Anthony Simeonov, Pulkit Agrawal, Abhishek Gupta

机构 * Massachusets Institute of Technology(麻省理工学院) University of Washington(华盛顿大学) Stanford University(斯坦福大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23044 2025-10-13 cs.CV 57%

SpatialSplat: Efficient Semantic 3D from Sparse Unposed Images

Yu Sheng, Jiajun Deng, Xinran Zhang, Yu Zhang, Bei Hua, Yanyong Zhang, Jianmin Ji

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07839 2025-10-10 cs.CV 57%

AlignGS: Aligning Geometry and Semantics for Robust Indoor Reconstruction from Sparse Views

Yijie Gao, Houqiang Zhong, Tianchi Zhu, Zhengxue Cheng, Qiang Hu, Li Song

机构 * School of Information Science and Electronic Engineering, Shanghai Jiao Tong University, Shanghai, China(信息科学与电子工程学院,上海交通大学,上海,中国) Cooperative Mediant Innovation Center, Shanghai Jiao Tong University, Shanghai, China(协同医疗创新中心,上海交通大学,上海,中国) SJTU Paris Elite Institute of Technology, Shanghai Jiao Tong University, Shanghai, China(上海交通大学巴黎精英技术学院)

专题命中 三维重建 :novel view synthesis(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07729 2025-10-10 cs.CV 57%

ComGS: Efficient 3D Object-Scene Composition via Surface Octahedral Probes

Jian Gao, Mengqi Yuan, Yifei Zeng, Chang Zeng, Zhihao Li, Zhenyu Chen, Weichao Qiu, Xiao-Xiao Long, Hao Zhu, Xun Cao, Yao Yao

机构 * Nanjing University(南京大学) Huawei Noah’s Ark Lab(华为诺亚实验室)

专题命中 三维重建 :Gaussian Splatting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03363 2025-10-09 cs.CV cs.AI eess.IV 57%

Unified Unsupervised Anomaly Detection via Matching Cost Filtering

Zhe Zhang, Mingxiu Cai, Gaochang Wu, Jing Zhang, Lingqiao Liu, Dacheng Tao, Tianyou Chai, Xiatian Zhu

机构 * State Key Laboratory of Synthetical Automation for Process Industries, Northeastern University, Shenyang, China(合成过程工业综合自动化国家重点实验室,东北大学,沈阳,中国) University of Surrey(Surrey大学) School of Computer Science, Wuhan University(武汉大学计算机学院) School of Computer Science, The University of Adelaide(阿德莱德大学计算机学院) College of Computing & Data Science, Nanyang Technological University(南洋理工大学计算机与数据科学学院) Surrey Institute for People-Centred Artificial Intelligence, and Centre for Vision, Speech and Signal Processing, University of Surrey(Surrey人本人工智能研究所,以及视觉、语音和信号处理中心,Surrey大学)

专题命中 三维重建 :point cloud(abstract);分类 cs.CV

Comments 63 pages (main paper and supplementary material), 39 figures, 58 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06621 2025-10-09 eess.IV cs.CE cs.CV cs.LG 57%

FEAorta: A Fully Automated Framework for Finite Element Analysis of the Aorta From 3D CT Images

Jiasong Chen, Linchen Qian, Ruonan Gong, Christina Sun, Tongran Qin, Thuy Pham, Caitlin Martin, Mohammad Zafar, John Elefteriades, Wei Sun, Liang Liang

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20297 2025-10-08 cs.RO 57%

mindmap: Spatial Memory in Deep Feature Maps for 3D Action Policies

Remo Steiner, Alexander Millane, David Tingdahl, Clemens Volk, Vikram Ramasamy, Xinjie Yao, Peter Du, Soha Pouya, Shiwei Sheng

机构 * NVIDIA(英伟达) Zurich Switzerland(苏黎世瑞士) Santa Clara California(圣克拉拉加州)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.RO

Comments Accepted to CoRL 2025 Workshop RemembeRL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05560 2025-10-08 cs.CV 57%

HoloScene: Simulation-Ready Interactive 3D Worlds from a Single Video

Hongchi Xia, Chih-Hao Lin, Hao-Yu Hsu, Quentin Leboutet, Katelyn Gao, Michael Paulitsch, Benjamin Ummenhofer, Shenlong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Intel(英特尔)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments Project page: https://xiahongchi.github.io/HoloScene

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04312 2025-10-07 cs.CV 57%

CARE-PD: A Multi-Site Anonymized Clinical Dataset for Parkinson's Disease Gait Assessment

Vida Adeli, Ivan Klabucar, Javad Rajabi, Benjamin Filtjens, Soroush Mehraban, Diwei Wang, Hyewon Seo, Trung-Hieu Hoang, Minh N. Do, Candice Muller, Claudia Oliveira, Daniel Boari Coelho, Pieter Ginis, Moran Gilat, Alice Nieuwboer, Joke Spildooren, Lucas Mckay, Hyeokhyen Kwon, Gari Clifford, Christine Esper, Stewart Factor, Imari Genias, Amirhossein Dadashzadeh, Leia Shum, Alan Whone, Majid Mirmehdi, Andrea Iaboni, Babak Taati

机构 * University of Toronto(多伦多大学) Vector Institute(向量研究所) KITE Research Institute-UHN(KITE研究 institute-UHN) University of Strasbourg(斯特拉斯堡大学) University Hospitals of Strasbourg(斯特拉斯堡大学医院) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Federal University of ABC(巴西ABC联邦大学) KU Leuven(鲁文大学) Hasselt University(哈瑟尔特大学) Emory University(埃默里大学) University of Bristol(布里斯托尔大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments Accepted at the Thirty-Ninth Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.19568 2025-10-07 cs.CV cs.AI 57%

How Far are AI-generated Videos from Simulating the 3D Visual World: A Learned 3D Evaluation Approach

Chirui Chang, Jiahui Liu, Zhengzhe Liu, Xiaoyang Lyu, Yi-Hua Huang, Xin Tao, Pengfei Wan, Di Zhang, Xiaojuan Qi

机构 * The University of Hong Kong(香港大学) Kling Team, Kuaishou Technology(快手科技 Kling 团队) Lingnan University(岭大)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03353 2025-10-07 cs.CV 57%

Sonar Image Datasets: A Comprehensive Survey of Resources, Challenges, and Applications

Larissa S. Gomes, Gustavo P. Almeida, Bryan U. Moreira, Marco Quiroz, Breno Xavier, Lucas Soares, Stephanie L. Brião, Felipe G. Oliveira, Paulo L. J. Drews-Jr

机构 * Instituto de Ciências Exatas e Tecnologia (ICET). Universidade Federal do Amazonas(巴西亚马逊联邦大学理学院(ICET))

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments Published in the Conference on Graphics, Patterns and Images (SIBGRAPI). This 4-page paper presents a timeline of publicly available datasets up to the year 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03198 2025-10-06 cs.CV 57%

Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft

Junchao Huang, Xinting Hu, Boyao Han, Shaoshuai Shi, Zhuotao Tian, Tianyu He, Li Jiang

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments 19 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17650 2025-10-06 cs.CV 57%

Evict3R: Training-Free Token Eviction for Memory-Bounded Streaming Visual Geometry Transformers

Soroush Mahdi, Fardin Ayar, Ehsan Javanmardi, Manabu Tsukada, Mahdi Javanmardi

机构 * Amirkabir University of Technology (AUT)(阿米尔卡比尔技术大学) The University of Tokyo(东京大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments project page: https://soroush-mim.github.io/projects/evict3r/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01183 2025-10-02 cs.CV 57%

EvoWorld: Evolving Panoramic World Generation with Explicit 3D Memory

Jiahao Wang, Luoxin Ye, TaiMing Lu, Junfei Xiao, Jiahan Zhang, Yuxiang Guo, Xijun Liu, Rama Chellappa, Cheng Peng, Alan Yuille, Jieneng Chen

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments Code available at: https://github.com/JiahaoPlus/EvoWorld

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10836 2025-10-02 cs.CV 57%

SL$^{2}$A-INR: Single-Layer Learnable Activation for Implicit Neural Representation

Moein Heidari, Reza Rezaeian, Reza Azad, Dorit Merhof, Hamid Soltanian-Zadeh, Ilker Hacihaliloglu

机构 * University of British Columbia(不列颠哥伦比亚大学) University of Tehran(塔里斯坦大学) RWTH Aachen University(亚琛工业大学) University of Regensburg(莱茵河畔大学)

专题命中 三维重建 :novel view synthesis(abstract);分类 cs.CV

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26498 2025-10-01 cs.CV 57%

DEPTHOR++: Robust Depth Enhancement from a Real-World Lightweight dToF and RGB Guidance

Jijun Xiang, Longliang Liu, Xuan Zhu, Xianqi Wang, Min Lin, Xin Yang

机构 * School of Electronic Information and Communications, Huazhong University of Science and Technology(电子信息与通信学院,华中科技大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments 15 pages, 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22527 2025-09-29 cs.CV 57%

EfficientDepth: A Fast and Detail-Preserving Monocular Depth Estimation Model

Andrii Litvynchuk, Ivan Livinsky, Anand Ravi, Nima Kalantari, Andrii Tsarov

机构 * Leia Inc.(Leia公司) Texas A&M University(德克萨斯农工大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments 12 pages, 7 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21859 2025-09-29 cs.CV 57%

SRHand: Super-Resolving Hand Images and 3D Shapes via View/Pose-aware Neural Image Representations and Explicit 3D Meshes

Minje Kim, Tae-Kyun Kim

机构 * School of Computing, KAIST(计算机学院,韩国科学技术院)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21664 2025-09-29 cs.RO cs.LG 57%

Generating Stable Placements via Physics-guided Diffusion Models

Philippe Nadeau, Miguel Rogel, Ivan Bilić, Ivan Petrović, Jonathan Kelly

机构 * STARS Laboratory, University of Toronto Institute for Aerospace Studies(多伦多大学航空航天研究所STARS实验室) Laboratory for Autonomous Systems and Mobile Robotics(自主系统与移动机器人实验室) University of Zagreb Faculty of Electrical Engineering and Computing(Zagreb大学电气工程与计算学院)

专题命中 三维重建 :point cloud(abstract);分类 cs.RO

Comments Submitted to the IEEE International Conference on Robotics and Automation 2026, Vienna, Austria, June 1-5, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏