arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-03-05 至 2026-03-05 共收录 16 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 6 篇

2602.24065 2026-03-05 cs.CV 79%

EvalMVX: A Unified Benchmarking for Neural 3D Reconstruction under Diverse Multiview Setups

EvalMVX: 一种用于神经3D重建的统一基准测试,适用于多样化的多视角设置

Zaiyan Yang, Jieji Ren, Xiangyi Wang, zonglin li, Xu Cao, Heng Guo, Zhanyu Ma, Boxin Shi

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) School of Mechanical Engineering, Shanghai Jiao Tong University(上海交通大学机械工程学院) Xiong’an Aerospace Information Research Institute(雄安航空航天信息研究所) State Key Laboratory of Multimedia Information Processing, School of Computer Science Peking University(北京大学计算机科学学院多媒体信息处理国家重点实验室)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 EvalMVX是一个包含25个物体的真实世界数据集,用于评估神经3D重建方法在不同多视角设置下的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03654 2026-03-05 cs.CV cs.AI eess.IV 70%

Field imaging framework for morphological characterization of aggregates with computer vision: Algorithms and applications

用于颗粒形态表征的现场成像框架:算法与应用

Haohang Huang

专题命中 三维重建 :3D reconstruction(abstract);point cloud(abstract);分类 cs.CV

AI总结 本研究提出了一种用于颗粒形态表征的现场成像框架,结合3D重建、分割和完成技术,实现对建筑用料的多场景高效分析与预测。

Comments PhD thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03950 2026-03-05 cs.CV cs.AI 70%

Improving Multi-View Reconstruction via Texture-Guided Gaussian-Mesh Joint Optimization

通过纹理引导的高斯-网格联合优化改进多视图重建

Zhejia Cai, Puhua Jiang, Shiwei Mao, Hongkun Cao, Ruqi Huang

机构 * SIGS, Tsinghua University(清华大学信息科学与技术学院) Peng Cheng Laboratory(鹏城实验室)

专题命中 三维重建 :3D reconstruction(abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 本文提出一种通过纹理引导的高斯-网格联合优化方法,实现多视图重建中几何与外观的统一优化,提升3D重建质量以支持后续编辑任务。

Comments 10 pages, correct errors, clarify details, accepted to 3DV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03099 2026-03-05 cs.RO 57%

Point2Act: Efficient 3D Distillation of Multimodal LLMs for Zero-Shot Context-Aware Grasping

Point2Act: 多模态大语言模型的高效3D蒸馏用于零样本情境感知抓取

Sang Min Kim, Hyeongjun Heo, Junho Kim, Yonghyeon Lee, Young Min Kim

机构 * Seoul National University(首尔国立大学) Massachusetts Institute of Technology(麻省理工学院)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.RO

AI总结 Point2Act通过多模态大语言模型高效蒸馏实现零样本情境感知抓取,生成空间定位响应以支持实际操作任务。

Comments Accepted to ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04368 2026-03-05 cs.NI 50%

LLM-supported 3D Modeling Tool for Radio Radiance Field Reconstruction

支持大型语言模型的3D建模工具用于无线电辐射场重建

Chengling Xu, Huiwen Zhang, Haijian Sun, Feng Ye

专题命中 三维重建 :3DGS(abstract)

AI总结 本文提出了一种结合微调语言模型和生成3D建模框架的本地部署工具,用于简化无线电辐射场重建的3D环境创建。

Comments Submitted to an IEEE conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03342 2026-03-05 eess.IV cs.AI q-bio.BM 50%

Cryo-SWAN: the Multi-Scale Wavelet-decomposition-inspired Autoencoder Network for molecular density representation of molecular volumes

Cryo-SWAN: 多尺度小波分解启发的自编码网络用于分子密度的分子体积表示

Rui Li, Artsemi Yushkevich, Mikhail Kudryashev, Artur Yakimovich

机构 * Center for Advanced Systems Understanding (CASUS)(先进系统理解中心(CASUS)) Helmholtz-Zentrum Dresden-Rossendorf e. V. (HZDR)(德累斯顿-罗斯托克亥姆霍尔茨研究中心(HZDR)) In situ Structural Biology(原位结构生物学) Department of Physics(物理系) Institute of Medical Physics and Biophysics(医学物理与生物物理研究所) Institute of Computer Science(计算机科学研究所) Cluster of Excellence Physics of Life(生命物理卓越中心)

专题命中 三维重建 :point cloud(abstract)

AI总结 Cryo-SWAN通过多尺度小波分解启发的自编码网络,实现分子密度体积的鲁棒表示,提升3D重建和生成质量。

Comments 16 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

2. Gaussian Splatting 3 篇

2603.04254 2026-03-05 cs.CV 85%

EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding

EmbodiedSplat: 在线前馈语义3DGS用于开放词汇3D场景理解

Seungjun Lee, Zihan Wang, Yunsong Wang, Gim Hee Lee

机构 * National University of Singapore(新加坡国立大学)

专题命中 Gaussian Splatting :3DGS(title,abstract);3D reconstruction(abstract);point cloud(abstract);分类 cs.CV

AI总结 EmbodiedSplat通过在线前馈3DGS实现开放词汇3D场景理解,结合CLIP编码和3D U-Net提升语义重建效率与泛化能力。

Comments CVPR 2026, Project Page: https://0nandon.github.io/EmbodiedSplat/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02887 2026-03-05 cs.GR cs.CV 84%

Generalized non-exponential Gaussian splatting

广义非指数高斯散射

Sébastien Speierer, Adrian Jarabo

机构 * Meta

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV、cs.GR

AI总结 本文将3DGS推广到非指数辐射传输模型,通过二次透射率定义不同衰减特性版本,提升复杂真实场景下的渲染效率。

Comments 13 pages, 6 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03602 2026-03-05 cs.CV 57%

DM-CFO: A Diffusion Model for Compositional 3D Tooth Generation with Collision-Free Optimization

DM-CFO: 一种用于无碰撞优化的组合3D牙齿生成扩散模型

Yan Tian, Pengcheng Xue, Weiping Ding, Mahmoud Hassaballah, Karen Egiazarian, Aura Conci, Abdulkadir Sengur, Leszek Rutkowski

机构 * School of Computer Science and Technology, Zhejiang Gongshang University(浙江工商大学计算机科学与技术学院) AGH University of Krakow(克拉科夫AGH大学) Polish Ministry of Science and Higher Education(波兰教育部) Tongxiang Institute of General Artificial Intelligence(同祥通用人工智能研究院) State Key Laboratory of Advanced Medical Materials and Devices(先进医学材料与器件国家重点实验室) School of Artificial Intelligence and Computer Science, Nantong University(南通大学人工智能与计算机科学学院) Department of Computer Science, College of Computer Engineering and Sciences, Prince Sattam Bin Abdulaziz University(普莱斯·本·阿卜杜勒阿齐兹大学计算机科学系) Department of Computer Science, Qena University(基纳大学计算机科学系) Department of Computing Sciences, Tampere University(塔尔库大学计算科学系) Department of Computer Science, Universidade Federal Fluminense(里约热内卢联邦大学计算机科学系) Shining3D Tech Co., Ltd.(Shining3D科技有限公司)

专题命中 Gaussian Splatting :3D generation(abstract);分类 cs.CV

AI总结 DM-CFO通过扩散模型和无碰撞优化生成组合3D牙齿,提升生成牙齿的多视图一致性和真实性。

Comments Received by IEEE Transactions on Visualization and Computer Graphics

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 点云 6 篇

2603.04208 2026-03-05 cs.RO 79%

GSeg3D: A High-Precision Grid-Based Algorithm for Safety-Critical Ground Segmentation in LiDAR Point Clouds

GSeg3D:一种高精度基于网格的算法,用于激光雷达点云中的安全关键地面分割

Muhammad Haider Khan Lodhi, Christoph Hertzberg

机构 * German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)

专题命中 点云 :point cloud(title,abstract);分类 cs.RO

AI总结 GSeg3D提出了一种高精度基于网格的地面分割算法,以满足自动驾驶和机器人在安全关键场景中的高精度需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04099 2026-03-05 cs.CV cs.AI 79%

Efficient Point Cloud Processing with High-Dimensional Positional Encoding and Non-Local MLPs

高效点云处理与高维位置编码及非局部MLP

Yanmei Zou, Hongshan Yu, Yaonan Wang, Zhengeng Yang, Xieyuanli Chen, Kailun Yang, Naveed Akhtar

机构 * School of Artificial Intelligence and Robotics, Quanzhou Institute of Industrial Design and Machine Intelligence Innovation, Hunan University(人工智能与机器人学院、泉州工业设计与智能机械创新研究院、湖南大学) College of Engineering and Design, Hunan Normal University(工程与设计学院、湖南师范大学) College of Intelligence Science and Technology, National University of Defense Technology(智能科学与技术学院、国防科技大学) School of Computing and Information Systems, The University of Melbourne(计算与信息学院、墨尔本大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出了一种基于高维位置编码和非局部MLP的点云处理方法,通过两阶段抽象和细化框架提升效率与效果。

Comments Accepted to IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI). Source code is available at https://github.com/zouyanmei/HPENet_v2.git

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14142 2026-03-05 cs.CV 79%

Token Adaptation via Side Graph Convolution for Efficient Fine-tuning of 3D Point Cloud Transformers

通过侧图卷积实现标记适应以实现3D点云变换器的高效微调

Takahiko Furuya

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 STAG通过侧图卷积实现高效微调,减少参数和计算成本,提升3D点云变换器性能。

Comments Accepted to the journal of Machine Vision and Applications

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03890 2026-03-05 eess.IV 78%

Point Cloud Feature Coding for Object Detection over an Error-Prone Cloud-Edge Collaborative System

点云特征编码用于错误多的云边协作系统中的目标检测

Chongzhen Tian, Hui Yuan, Pan Zhao, Chang Sun, Raouf Hamzaoui, Sam Kwong

专题命中 点云 :point cloud(title,abstract)

AI总结 本文提出了一种基于源编码和信道编码的任务驱动点云压缩和可靠传输框架,用于在错误多的云边协作系统中实现高效的目标检测。

Comments 13 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03960 2026-03-05 cs.RO cs.CV 62%

Structural Action Transformer for 3D Dexterous Manipulation

结构动作变换器用于3D灵巧操作

Xiaohan Lei, Min Wang, Bohong Weng, Wengang Zhou, Houqiang Li

机构 * MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知国家重点实验室,中国科学技术大学) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究院,合肥综合性国家科学中心)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 本文提出结构动作变换器,通过结构化视角解决高自由度机械手的跨身体技能转移问题,实现更高效的灵巧操作。

Comments Accepted by CVPR

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17896 2026-03-05 cs.CV cs.AI 57%

EgoWorld: Translating Exocentric View to Egocentric View using Rich Exocentric Observations

EgoWorld:利用丰富的外参照观测将外参照视角转换为自身参照视角

Junho Park, Andrew Sangwoo Ye, Taein Kwon

机构 * AI Lab, LG Electronics(LG电子人工智能实验室) KAIST(韩国科学技术院) Visual Geometry Group, University of Oxford(牛津大学视觉几何组)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 EgoWorld通过重建丰富的外参照观测来实现外参照到自身参照视角的转换,展示了在增强现实、虚拟现实和机器人应用中的先进性能和鲁棒性。

Comments Accepted by ICLR 2026. Project Page: https://redorangeyellowy.github.io/EgoWorld/

详情

展开后加载摘要…

URL PDF HTML 收藏

4. SLAM与定位 1 篇

2603.03453 2026-03-05 cs.RO 57%

Radar-based Pose Optimization for HD Map Generation from Noisy Multi-Drive Vehicle Fleet Data

基于雷达的姿态优化用于从噪声多车车队数据生成HD地图

Alexander Blumberg, Jonas Merkert, Christoph Stiller

机构 * Institute of Measurement and Control Systems, Karlsruhe Institute of Technology (KIT)(测量与控制系统研究所,卡尔斯鲁厄理工学院)

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.RO

AI总结 本文提出基于雷达的姿态优化方法,用于从噪声多车车队数据生成高精度地图,通过姿态图优化提升地图精度和特征清晰度。

Comments Accepted for the 37th IEEE Intelligent Vehicles Symposium (IV 2026), 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏