arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-03-18 至 2026-03-18 共收录 23 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 4 篇

2603.16133 2026-03-18 cs.CV 79%

DualPrim: Compact 3D Reconstruction with Positive and Negative Primitives

DualPrim: 基于正负原语的紧凑3D重建

Xiaoxu Meng, Zhongmin Chen, Bo Yang, Weikai Chen, Weixiao Liu, Lin Gao

机构 * Independent Researcher(独立研究者) Waymo LLC(Waymo公司) Lucid Motors Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) School of Advanced Interdisciplinary Science, University of Chinese Academy of Sciences(中国科学院大学先进交叉学科学院)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 DualPrim通过正负超球体实现紧凑且结构化的3D重建,提升拓扑感知建模能力,实现端到端学习和无缝导出。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11508 2026-03-18 cs.CV 79%

On Geometric Understanding and Learned Priors in Feed-forward 3D Reconstruction Models

关于馈送式3D重建模型中的几何理解与学习先验

Jelena Bratulić, Sudhanshu Mittal, Thomas Brox, Christian Rupprecht

机构 * University of Freiburg(弗赖堡大学) University of Oxford(牛津大学)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 研究探讨了馈送式3D重建模型是否基于传统多视图流程的几何原理或依赖大规模训练的学习先验,通过分析内部表示和注意力模式验证了几何理解的存在。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16612 2026-03-18 cs.GR 57%

Retrieval-Augmented Sketch-Guided 3D Building Generation

增强检索的草图引导3D建筑生成

Zhengyang Wang, Nuttapong Rochanavibhata, Yuxiao Ren, Xusheng Du, Ye Zhang, Haoran Xie

专题命中 三维重建 :3D generation(abstract);分类 cs.GR

AI总结 本文提出多阶段3D生成框架,通过结合生成与检索方法实现组件级编辑与个性化定制,解决日本独栋房屋早期设计阶段因缺乏统一表示导致的设计漂移问题。

Comments 10 pages, 4 figures, Proceeding of CAADRIA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07107 2026-03-18 cs.CV 57%

COREA: Coupled Relightable 3D Gaussians and SDFs for Efficient Normal Alignment

COREA:耦合可重照明3D高斯与SDF用于高效的法线对齐

Jaeyoon Lee, Hojoon Jung, Sungtae Hwang, Jihyong Oh, Jongwon Choi

机构 * Chung-Ang University(Chung-Ang 大学)

专题命中 三维重建 :3DGS(abstract);分类 cs.CV

AI总结 COREA是首个统一的三任务框架,结合SDF和可重照明3D高斯(3DGS)以联合支持基于球谐函数的新型视角合成(NVS)、表面重建和逆物理渲染(逆PBR)。通过共享底层表面,利用几何约束的可重照明3DGS提供可靠深度信号,SDF的连续法线场为高斯法线学习提供空间一致监督。

Comments Project page: https://cau-vilab.github.io/COREA/

详情

展开后加载摘要…

URL PDF HTML 收藏

2. NeRF 1 篇

2603.15622 2026-03-18 cs.CV cs.AI 83%

SAC-NeRF: Adaptive Ray Sampling for Neural Radiance Fields via Soft Actor-Critic Reinforcement Learning

SAC-NeRF:通过软演员-评论员强化学习实现神经辐射场的自适应射线采样

Chenyu Ge

机构 * University of Southern California(南加州大学)

专题命中 NeRF :NeRF(title,abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 本文提出SAC-NeRF,通过强化学习框架学习自适应采样策略,减少采样点35-48%,保持渲染质量与密集采样基线相近。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. Gaussian Splatting 9 篇

2603.16211 2026-03-18 cs.CV 90%

Leveling3D: Leveling Up 3D Reconstruction with Feed-Forward 3D Gaussian Splatting and Geometry-Aware Generation

Leveling3D: 通过前馈3D高斯点划法与几何感知生成提升3D重建

Yiming Huang, Baixiang Huang, Beilei Cui, Chi Kit Ng, Long Bai, Hongliang Ren

机构 * The Chinese University of Hong Kong(香港中文大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3D reconstruction(title,abstract);3D vision(abstract);3DGS(abstract)

AI总结 Leveling3D结合前馈3D重建与几何一致生成,解决传统方法在 extrapolated view 中的缺失区域问题,通过几何感知适配器提升3D重建质量,实现生成与重建的同步优化。

Comments 26 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17531 2026-03-18 cs.GR cs.AI cs.CV 86%

Laplace-Beltrami Operator for Gaussian Splatting

高斯散射中的拉普拉斯-贝尔特拉米算子

Hongyu Zhou, Zorah Lähner

机构 * University of Bonn(波恩大学) Lamarr Institute for Machine Learning and Artificial Intelligence(拉马尔人工智能与机器学习研究院)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3D reconstruction(abstract);point cloud(abstract);分类 cs.CV、cs.GR

AI总结 本文提出在高斯散射表示上直接计算拉普拉斯-贝尔特拉米算子的方法,利用马哈拉诺斯距离提高几何处理的准确性,并用于评估优化结果的质量。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16538 2026-03-18 cs.CV 83%

Rethinking Pose Refinement in 3D Gaussian Splatting under Pose Prior and Geometric Uncertainty

重新思考在姿态先验和几何不确定性下的3D高斯点散布姿态细化

Mangyu Kong, Jaewon Lee, Seongwon Lee, Euntai Kim

机构 * Yonsei University(延世大学) Kookmin University(韩国庆熙大学) Korea Institution of Science and Technology(韩国科学技术院)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 本文针对3DGS姿态细化的鲁棒性问题,提出结合蒙特卡洛姿态采样与Fisher信息PnP优化的重定位框架,提升定位精度和稳定性。

Comments 17 pages, 11 figures, CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16844 2026-03-18 cs.CV 79%

M^3: Dense Matching Meets Multi-View Foundation Models for Monocular Gaussian Splatting SLAM

M^3:密集匹配与多视角基础模型结合的单目高斯点云SLAM

Kerui Ren, Guanghao Li, Changjian Jiang, Yingxiang Xu, Tao Lu, Linning Xu, Junting Dong, Jiangmiao Pang, Mulin Yu, Bo Dai

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Fudan University(复旦大学) Shanghai Innovation Institute(上海创新研究院) Zhejiang University(浙江大学) Beijing Institute of Technology(北京理工大学) The Chinese University of Hong Kong(香港中文大学) The University of Hong Kong(香港大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 本文提出M^3,结合多视角基础模型与单目高斯点云SLAM,通过引入匹配头提升密集对应关系,增强跟踪稳定性,在多种基准上取得优异的位姿估计和场景重建性能。

Comments Project page: https://city-super.github.io/M3/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16154 2026-03-18 cs.CV cs.AI 79%

GATS: Gaussian Aware Temporal Scaling Transformer for Invariant 4D Spatio-Temporal Point Cloud Representation

GATS: 基于高斯意识的时间缩放变换器的不变四维时空点云表示

Jiayi Tian, Jiaze Wang

机构 * State Key Laboratory of Human-Machine Hybrid Augmented Intelligence(人机混合增强智能国家重点实验室) Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院) Xi’an Jiaotong University(西安交通大学) Harbin Institute of Technology(哈尔滨工业大学) Pengcheng Laboratory(鹏城实验室)

专题命中 Gaussian Splatting :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出GATS框架,通过不确定性引导的高斯卷积和时间缩放注意力模块,解决四维点云视频中时间尺度偏差和分布不确定性问题,提升鲁棒性和精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09881 2026-03-18 cs.CV 70%

LTGS: Long-Term Gaussian Scene Chronology From Sparse View Updates

LTGS: 从稀疏视角更新中获取长期高斯场景时间线

Minkwan Kim, Seungmin Lee, Junho Kim, Young Min Kim

机构 * Dept. of Electrical and Computer Engineering, Seoul National University(电子与计算机工程系,首尔国立大学) Interdisciplinary Program in Artificial Intelligence and INMC, Seoul National University(人工智能跨学科项目及INMC,首尔国立大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3DGS(abstract);分类 cs.CV

AI总结 本文提出LTGS方法,通过稀疏视角更新高效表示日常场景变化,利用模板高斯作为结构先验,提升3D环境的时间演化可扩展性。

Comments Accepted to CVPR 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13911 2026-03-18 cs.CV 70%

PhysGM: Large Physical Gaussian Model for Feed-Forward 4D Synthesis

PhysGM:用于前馈4D合成的大型物理高斯模型

Chunji Lv, Zequn Chen, Donglin Di, Weinan Zhang, Hao Li, Wei Chen, Yinjie Lei, Changsheng Li

机构 * Beijing Institute of Technology(北京理工大学) Li Auto(利汽车) Harbin Institute of Technology(哈尔滨工业大学) Sichuan University(四川大学) Suzhou Research Institute, Harbin Institute of Technology(哈尔滨工业大学苏州研究院)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3DGS(abstract);分类 cs.CV

AI总结 PhysGM通过单图像联合预测3D高斯表示和物理属性,实现高效4D渲染。该方法采用物理感知重建模型和直接偏好优化,提升模拟精度并降低计算成本。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06510 2026-03-18 eess.IV 67%

Three-Dimensional MRI Reconstruction with Gaussian Representations: Tackling the Undersampling Problem

三维MRI重建与高斯表示:解决欠采样问题

Tengya Peng, Ruyi Zha, Zhen Li, Xiaofeng Liu, Qing Zou

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3DGS(abstract)

AI总结 本文提出3D高斯MRI框架,利用高斯分布重建等效分辨率的3D MRI,无需大量数据集,有效提升重建质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13639 2026-03-18 cs.RO 57%

4D Radar-Inertial Odometry based on Gaussian Modeling and Multi-Hypothesis Scan Matching

基于高斯建模和多假设扫描匹配的4D雷达-惯性里程计

Fernando Amodeo, Luis Merino, Fernando Caballero

机构 * Service Robotics Laboratory, Universidad Pablo de Olavide(服务机器人实验室, Pablo de Olavide 大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.RO

AI总结 本文提出基于高斯分布的4D雷达场景表示方法,通过全局优化提升注册精度,并通过多假设优化避免局部最优,实验表明其在雷达-惯性里程计任务中表现优异。

Comments Our code and results can be publicly accessed at: https://github.com/robotics-upo/gaussian-rio-cpp Accepted for publication in IEEE Robotics and Automation Letters

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 点云 7 篇

2603.16781 2026-03-18 cs.CV cs.AI 79%

IOSVLM: A 3D Vision-Language Model for Unified Dental Diagnosis from Intraoral Scans

IOSVLM:一种用于从牙科内窥扫描进行统一牙科诊断的3D视觉-语言模型

Huimin Xiong, Zijie Meng, Tianxiang Hu, Chenyi Zhou, Yang Feng, Zuozhu Liu

机构 * ZJU-UIUC Institute, Zhejiang University, Haining, 314400, China(浙大浙ICU研究所,浙江大学,海宁,314400,中国) Stomatology Hospital, School of Stomatology, Zhejiang University School of Medicine, Hangzhou, 310058, China(口腔医院,口腔医学院,浙江大学医学院,杭州,310058,中国) Angelalign Research Institute, Angel Align Inc., Shanghai, 200011, China(天使对齐研究院,天使对齐公司,上海,200011,中国)

专题命中 点云 :3D vision(title);point cloud(abstract);分类 cs.CV

AI总结 本文提出IOSVLM,一种端到端的3D视觉-语言模型,利用点云表示内窥扫描,并结合大规模多源数据集,提升牙科诊断和生成式视觉问答的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16343 2026-03-18 cs.CV 79%

Learning Human-Object Interaction for 3D Human Pose Estimation from LiDAR Point Clouds

从LiDAR点云学习人类-物体交互以实现3D人体姿态估计

Daniel Sungho Jung, Dohee Cho, Kyoung Mu Lee

机构 * IPAI Dept. of ECE&ASRI(电子与信息科学研究院) Seoul National University(首尔国立大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出HOIL框架,通过学习人类-物体交互缓解LiDAR点云中3D人体姿态估计的空域模糊和类别不平衡问题。

Comments Project page: https://hoil-release.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15683 2026-03-18 stat.ML cs.LG 78%

Beyond Distance: Quantifying Point Cloud Dynamics with Persistent Homology and Dynamic Optimal Transport

超越距离:利用持久同调和动态最优传输量化点云动态

Yixin Wang, Ting Gao, Jinqiao Duan

机构 * Department of Mathematics and Department of Physics, Great Bay University, Dongguan, China(数学系和物理系,大湾大学,东莞,中国) Guangdong Provincial Key Laboratory of Mathematical and Neural Dynamical Systems, Dongguan, China(广东省数学与神经动力系统重点实验室,东莞,中国)

专题命中 点云 :point cloud(title,abstract)

AI总结 本文提出一种框架,通过扩展最近提出的拓扑最优传输(TpOT)距离,分析时间演化点云中的拓扑突变。引入分层动态评估框架,结合拓扑和超图重建策略,通过多尺度指标检测局部重构,验证了运输对齐与多尺度熵诊断结合在动态拓扑分析中的有效性。

Comments 42 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.14001 2026-03-18 cs.RO cs.CV 62%

CLAIM: Camera-LiDAR Alignment with Intensity and Monodepth

CLAIM: 相机-激光雷达对齐与强度和单目深度

Zhuo Zhang, Yonghui Liu, Meijie Zhang, Feiyang Tan, Yikang Ding

机构 * Mach Drive(马奇驱动)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 本文提出CLAIM方法,利用单目深度模型实现相机与激光雷达对齐,通过粗到细搜索优化结构损失和纹理损失,无需复杂数据处理,实验验证其优越性能。

Comments Accepted by IROS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19142 2026-03-18 cs.RO 57%

BiGraspFormer: End-to-End Bimanual Grasp Transformer

BiGraspFormer:端到端双臂抓取Transformer

Kangmin Kim, Seunghyeok Back, Geonhyup Lee, Sangbeom Lee, Sangjun Noh, Kyoobin Lee

机构 * Department of AI Convergence, Gwangju Institute of Science and Technology (GIST)(人工智能融合系,全州科学技术学院(GIST)) Department of AI Machinery, Korea Institute of Machinery & Materials (KIMM)(人工智能机械系,韩国机械材料研究院(KIMM))

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出BiGraspFormer,一种端到端的双臂抓取Transformer框架,通过单引导双臂策略生成协调的双臂抓取方案,有效解决现有方法在协调性和效率上的不足。

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16788 2026-03-18 eess.IV 50%

Preserving Vertical Structure in 3D-to-2D Projection for Permafrost Thaw Mapping

在3D到2D投影中保持垂直结构以进行永久冻土融化映射

Justin McMillen, Robert Van Alphen, Taha Sadeghi Chorsi, Jason Shabaga, Mel Rodgers, Rocco Malservisi, Timothy Dixon, Yasin Yilmaz

专题命中 点云 :point cloud(abstract)

AI总结 本文提出一种投影解码器,通过学习高度嵌入实现高度依赖的特征转换,结合分层采样保持垂直信息,用于预测永久冻土融化深度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16327 2026-03-18 math.AT 50%

Worst-Case Examples for the Computation of Persistent Homology

持久同调计算的最坏情况示例

Uzay Çetin, Ergun Yalcin

专题命中 点云 :point cloud(abstract)

AI总结 本文构造了持久同调标准简化算法的最坏情况示例,通过替换单三角形排列为基底和尖三角形带结构,提出了构造算法并进行算法运行时间实验,证明经过适当边和三角形细分后,这些示例仍为最坏情况,并可作为过滤图的克利克复形和有限点云的维特里希-里斯复形。

Comments 19 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 新视角合成 1 篇

2603.16306 2026-03-18 cs.CV 57%

DriveFix: Spatio-Temporally Coherent Driving Scene Restoration

DriveFix:时空一致的驾驶场景修复

Heyu Si, Brandon James Denis, Muyang Sun, Dragos Datcu, Yaoru Li, Xin Jin, Ruiju Fu, Yuliia Tatarinova, Federico Landi, Jie Song, Mingli Song, Qi Guo

机构 * Zhejiang University(浙江大学) Huawei(华为)

专题命中 新视角合成 :novel view synthesis(abstract);分类 cs.CV

AI总结 DriveFix提出一种多视角修复框架,通过交错扩散变换器架构和几何感知损失,实现驾驶场景时空一致性,提升4D世界建模的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 空间理解 1 篇

2411.16253 2026-03-18 cs.CV 57%

Open-Vocabulary Octree-Graph for 3D Scene Understanding

开放词汇八叉树图用于3D场景理解

Zhigang Wang, Yifei Su, Chenhui Li, Dong Wang, Yan Huang, Bin Zhao, Xuelong Li

机构 * Northwestern Polytechnical University(西北工业大学) Shanghai AI Laboratory(上海人工智能实验室) University of Chinese Academy of Sciences(中国科学院大学) CASIA TeleAI

专题命中 空间理解 :point cloud(abstract);分类 cs.CV

AI总结 本文提出Octree-Graph,通过CGSM和IFA算法获取3D实例及语义特征,构建适应性八叉树结构以高效表示场景,实验表明其在多种任务中具有广泛适用性和有效性。

Comments Accepted by ICCV25. 11 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏