arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-06-10 至 2026-06-10 共收录 16 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 2 篇

2606.10478 2026-06-10 cs.CV 新提交 87%

3D-CoS: A New 3D Reconstruction Paradigm Based on VLM Code Synthesis

3D-CoS:基于VLM代码合成的新型3D重建范式

Yuhao Wang, Puyi Wang, Linjie Li, Zhengyuan Yang, Kevin Qinghong Lin, Yu Cheng

机构 * Shanghai Jiao Tong University(上海交通大学) The Chinese University of Hong Kong(香港中文大学) Microsoft(微软) University of Oxford(牛津大学)

专题命中 三维重建 :3D reconstruction(title,abstract);NeRF(abstract,abstract_cn);point cloud(abstract);分类 cs.CV

AI总结 提出3D代码合成(3D-CoS)范式,将3D资产表示为可执行的Blender代码,利用VLM进行程序化重建,实现高可控性和局部编辑能力。

Comments Preprint. 24 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10142 2026-06-10 cs.CV 新提交 57%

DB-3DME: From Dataset to Benchmark for Human-aligned Automatic 3D Mesh Evaluation

DB-3DME:从数据集到基准测试,实现与人类对齐的自动3D网格评估

Nanshan Jia, Zhenyu Zhao, Sui Huang, Jingshen Wang, Zeyu Zheng

机构 * University of California, Berkeley(加州大学伯克利分校) Roblox Corporation(Roblox公司)

专题命中 三维重建 :3D generation(abstract);分类 cs.CV

AI总结 提出DB-3DME数据集与基准,包含2619个合成3D网格及其人类评分,通过微调视觉编码器优化VLM评估性能,显著超越现有模型。

Comments CVPR 2026 workshop paper. 10 pages, 3 figures, 6 tables. Dataset available at GitHub and Hugging Face

详情

展开后加载摘要…

URL PDF HTML 收藏

2. NeRF 1 篇

2508.18540 2026-06-10 cs.GR eess.IV 版本更新 57%

Real-time 3D Visualization of Radiance Fields on Light Field Displays

光场显示上辐射场的实时3D可视化

Jonghyun Kim, Cheng Sun, Michael Stengel, Matthew Chan, Andrew Russell, Jaehyun Jung, Wil Braithwaite, David Luebke, Shalini De Mello

专题命中 NeRF :Gaussian Splatting(abstract);分类 cs.GR

AI总结 针对光场显示需要多视角高分辨率渲染而辐射场计算密集的问题,提出统一高效框架,通过共享中间扫描平面单次合成密集光场视图,实现实时渲染(200+ FPS),速度提升22倍。

Comments 19 pages, 14 figures. J. Kim, C. Sun, and M. Stengel contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏

3. Gaussian Splatting 6 篇

2407.09510 2026-06-10 cs.CV 93%

3DGS.zip: A survey on 3D Gaussian Splatting Compression Methods

3DGS.zip:3D高斯散射压缩方法综述

Milena T. Bagdasarian, Paul Knoll, Yi-Hsin Li, Florian Barthel, Anna Hilsmann, Peter Eisert, Wieland Morgenstern

机构 * Fraunhofer HHI(弗劳恩霍夫研究所汉诺威研究所) Humboldt-Universität zu Berlin(柏林洪堡大学) Technische Universität Berlin(柏林技术大学)

专题命中 Gaussian Splatting :3DGS(title,title_cn);Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 本文综述了3DGS压缩方法,探讨了压缩与紧缩技术,旨在提高3DGS的效率和实用性,通过减少文件大小和高斯数量来优化质量和性能。

Comments 3D Gaussian Splatting compression survey; 3DGS compression; updated discussion; new approaches added; new illustrations

Journal ref Computer Graphics Forum, Volume 44, Issue 2 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10612 2026-06-10 cs.CV 新提交 85%

GaussTrace: Provenance Analysis of 3D Gaussian Splatting Models with Evidence-based LLM Reasoning

GaussTrace:基于证据的LLM推理的3D高斯泼溅模型溯源分析

Haoliang Han, Ziyuan Luo, Renjie Wan

机构 * University of California, Berkeley(加州大学伯克利分校) Stanford University(斯坦福大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract,abstract_cn);分类 cs.CV

AI总结 提出GaussTrace框架,通过属性统计分析和假设驱动的编辑模拟,结合大语言模型链式推理,构建3D高斯泼溅模型的有向溯源图,无需训练或编辑历史。

Comments Accepted by ICML2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09967 2026-06-10 cs.CV 新提交 81%

ABot-Earth 0.5: Generative 3D Earth Model

ABot-Earth 0.5:生成式3D地球模型

Ming Qian, Tianjian Ouyang, Mingchao Sun, Zijian Wang, Jincheng Xiong, Jiarong Han, Yongchang Zhang, Jiawei Zhang, Xu Wang, Yu Liu, Luyang Tang, Fei Yu, Zengye Ge, Mengmeng Du, Yuan Liu, Nianfei Fan, Song Wang, Yingliang Peng, Chunxue Jia, Yang Liu, Shiying Zeng, Haozhe Shi, Junnan Lai, Hongyu Pan, Zheng Wu, Ning Guo, Mu Xu, Hang Zhang

机构 * AMAP CV Lab(AMAP视觉实验室)

专题命中 Gaussian Splatting :3DGS(abstract,abstract_cn);Gaussian Splatting(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 提出ABot-Earth 0.5框架,利用3D高斯泼溅从卫星图像生成大规模无缝3D环境,每平方公里合成时间低于10分钟,支持实时交互可视化,降低3D重建成本。

Comments From Amap-cvlab, Alibaba. Official page: https://abot-earth.amap.com/

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10645 2026-06-10 cs.CV 新提交 79%

ManiSplat: Manipulation Trajectory Synthesis from Monocular Video via Decoupled 3D Gaussian Splatting

ManiSplat: 基于解耦3D高斯泼溅的单目视频操作轨迹合成

Wenhao Hu, Haonan Zhou, Liu Liu, Yun Du, Xinjie Wang, Ziang Li, Zhizhong Su, Gaoang Wang

机构 * Zhejiang University(浙江大学) Horizon Robotics(地平线机器人)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 提出ManiSplat框架,通过图结构解耦表示和任务导向时空对齐,从单目视频重建可控的3D高斯数字孪生,支持机器人操作任务与策略学习。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.19843 2026-06-10 cs.GR 版本更新 79%

Graphical X Splatting (GraphiXS): A Graphical Model for 4D Gaussian Splatting under Uncertainty

图形化X溅射(GraphiXS):不确定性下的4D高斯溅射图形模型

Doğa Yılmaz, Jialin Zhu, Deshan Gong, He Wang

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.GR

AI总结 提出GraphiXS概率框架,通过图形模型系统整合多种数据不确定性,扩展4D高斯溅射至概率设置,支持多种基元,在数据缺失或污染场景中优于现有方法。

Comments Accepted to SIGGRAPH 2026

Journal ref SIGGRAPH 2026 Conf. Proc

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10656 2026-06-10 cs.CV 新提交 74%

Envision4D: Envisioning Visual Futures via Feed-forward 4D Gaussian Splatting for Autonomous Driving

Envision4D: 通过前馈4D高斯泼溅展望自动驾驶的视觉未来

Qi Song, Yifei He, Chi Zhang, Zheng Fu, Xuhe Zhao, Mengmeng Yang, Kun Jiang, Rui Huang, Diange Yang

机构 * Tsinghua University(清华大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

专题命中 Gaussian Splatting :Gaussian Splatting(title);分类 cs.CV

AI总结 提出Envision4D,一种全自监督前馈框架,通过未来姿态预测、层内时间注意力和条件运动提升,实现无位姿的未来外推,在自动驾驶动态场景预测中达到最先进性能。

Comments Project Page: https://maggiesong7.github.io/research/Envision4D/

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 点云 7 篇

2606.10019 2026-06-10 cs.CV cs.AI cs.RO 新提交 81%

Generalized-CVO: Fast and Correspondence-Free Local Point Cloud Registration with Second Order Riemannian Optimization

广义CVO:基于二阶黎曼优化的快速无对应局部点云配准

Ray Zhang, Marcus Greiff, Thomas Lew, John Subosits

机构 * Toyota Research Institute(丰田研究院)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV、cs.RO

AI总结 提出一种基于几何表面结构和再生核希尔伯特空间嵌入的无对应局部点云配准方法,采用二阶流形优化实现高达10倍加速,在LiDAR和RGB-D跟踪及物体配准中显著降低漂移并提升鲁棒性。

Comments 16 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10541 2026-06-10 cs.CV 新提交 79%

GRAR: Glass-induced Reflection Artifact Removal in LiDAR Point Clouds

GRAR: LiDAR点云中玻璃引起的反射伪影去除

Wanpeng Shao, Zeyi Guo, Bo Zhang, Yifei Xue, Tie Ji, Yizhen Lao

机构 * College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院) School of Design, Hunan University(湖南大学设计学院) School of Finance and Statistics, Hunan University(湖南大学金融与统计学院)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 提出两阶段框架,先利用多模态视觉基础模型和几何线索精确分割玻璃区域,再基于物理驱动的反射感知局部-全局几何相似性描述符去除反射伪影,在多个公开数据集上优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10395 2026-06-10 cs.CV 新提交 79%

Efficient RWKV-based Representation Learning for 3D Point Clouds

基于高效RWKV的三维点云表示学习

Yun Liu, Xuefeng Yan, Liangliang Nan, Xianzhi Li, Peng Li, Zhe Zhu, Honghua Chen, Mingqiang Wei

机构 * School of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院) Shenzhen Institute of Research, Nanjing University of Aeronautics and Astronautics(南京航空航天大学深圳研究院) Collaborative Innovation Center of Novel Software Technology and Industrialization(新型软件技术与产业化协同创新中心) Urban Data Science section, Delft University of Technology(代尔夫特理工大学城市数据科学部) Huazhong University of Science and Technology(华中科技大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 提出P-RWKV模块,通过局部感知扩展和空间上下文增强,将RWKV从序列建模适配到3D点云,实现线性复杂度的全局依赖建模,在多项任务中以更低计算成本取得竞争性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10594 2026-06-10 cs.CV 新提交 61%

Segment and Select: Vision-Language Segmentation in 3D Scenarios

Segment and Select: 3D场景中的视觉-语言分割

Yulin Chen, Zhihang Zhong, Yuenan Hou

机构 * Shanghai AI Laboratory(上海人工智能实验室) University of Science and Technology of China(中国科学技术大学) Shanghai Jiaotong University(上海交通大学)

专题命中 点云 :3D vision(abstract,comments);分类 cs.CV

AI总结 提出SEGA3D范式,通过掩码候选生成器、大语言模型和语义空间选择器实现3D场景中基于语言指令的细粒度分割,在ScanNet和Matterport3D上分别提升8.3和5.3 mIoU。

Comments The core idea is to reformulate 3D vision-language segmentation as the segment-and-select paradigm (free from the superpoint dependency)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09882 2026-06-10 cs.CV cs.LG 新提交 57%

WHU-Infra3D: A Full-stack Multi-modal Dataset and Benchmark for 3D Roadside Infrastructure Inventory

WHU-Infra3D:面向3D路边基础设施清单的全栈多模态数据集与基准

Chong Liu, Luxuan Fu, Xuyu Feng, Zhen Dong, Bisheng Yang

机构 * State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing (LIESMARS)(信息工程测绘遥感国家重点实验室) Wuhan University(武汉大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 提出WHU-Infra3D多模态基准数据集,覆盖三城市53.8公里,融合全景图像与LiDAR点云,提供2D-3D实例关联和跨帧跟踪,支持基础设施状态诊断与属性识别,填补自动化维护数据集空白。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13595 2026-06-10 cs.CV 版本更新 57%

NoiseSDF2NoiseSDF: Learning Clean Neural Fields from Noisy Supervision

NoiseSDF2NoiseSDF: 从含噪监督中学习干净的神经场

Tengkai Wang, Weihao Li, Ruikai Cui, Shi Qiu, Nick Barnes

机构 * University of Cambridge(剑桥大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 提出NoiseSDF2NoiseSDF方法,通过最小化含噪SDF表示之间的MSE损失,从含噪点云中学习干净的神经SDF,实现隐式去噪和表面优化。

Comments 16 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10978 2026-06-10 physics.optics 新提交 50%

Event-based Scheimpflug LiDAR for Ultra-Fast Laser-Scanned Rangefinding

基于事件的Scheimpflug激光雷达用于超快激光扫描测距

Nathan Meraz, Alisha Whitehead, Suet Ying Chan, Ronan Taneja, Gabriella Mayrend, Joseph L. Greene

专题命中 点云 :point cloud(abstract)

AI总结 提出eSCHORTY系统,将事件传感器与调制连续波线激光结合,通过Scheimpflug几何实现每秒百万兆事件级密集3D点云,解决帧率限制和背景干扰问题。

Comments 19 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏