arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-05-01 至 2026-05-01 共收录 19 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 5 篇

2604.28064 2026-05-01 cs.CV 83%

3D Reconstruction Techniques in the Manufacturing Domain: Applications, Research Opportunities and Use Cases

制造领域中的3D重建技术:应用、研究机会与用例

Chialoon Cheng, Kaijun liu, Zhiyang Liu, Marcelo H Ang

机构 * Advanced Robotics Centre(先进机器人中心) National University of Singapore(新加坡国立大学) Independent Researcher(独立研究员)

专题命中 三维重建 :3D reconstruction(title,abstract);point cloud(abstract);分类 cs.CV

AI总结 本文综述了制造领域3D重建技术的发展现状,分析了传统方法与深度学习方法的差异,指出统一框架的不足,并总结了质量检测等关键应用及未来研究方向。

Comments 24 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.23065 2026-05-01 cs.GR cs.CV 82%

EAG-PT: Emission-Aware Gaussians and Path Tracing for Diffuse Indoor Scene Reconstruction and Editing

EAG-PT:考虑发射的高斯和路径追踪用于漫反射室内场景重建与编辑

Xijie Yang, Mulin Yu, Changjian Jiang, Kerui Ren, Tao Lu, Jiangmiao Pang, Dahua Lin, Bo Dai, Linning Xu

机构 * Zhejiang University(浙江大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Shanghai Jiao Tong University(上海交通大学) The Chinese University of Hong Kong(香港中文大学) The University of Hong Kong(香港大学)

专题命中 三维重建 :NeRF(abstract,abstract_cn);3DGS(abstract,abstract_cn);分类 cs.CV、cs.GR

AI总结 EAG-PT通过统一的2D高斯表示实现室内场景的物理重建与渲染,支持可编辑的漫反射全局光照,结合高效单次反弹优化和高质量多次反弹路径追踪,优于现有方法。

Comments SIGGRAPH 2026 Conference Paper; Project Page: https://eag-pt.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.28193 2026-05-01 cs.CV 74%

Generalizable Sparse-View 3D Reconstruction from Unconstrained Images

可泛化稀疏视角3D重建从无约束图像

Vinayak Gupta, Chih-Hao Lin, Shenlong Wang, Anand Bhattad, Jia-Bin Huang

机构 * University of Maryland, College Park(马里兰大学学院公园分校) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Johns Hopkins University(约翰霍普金斯大学)

专题命中 三维重建 :3D reconstruction(title);分类 cs.CV

AI总结 本文提出GenWildSplat框架,通过学习几何先验在无场景优化情况下实现稀疏视角户外3D重建,利用外观适配器和语义分割处理光照和遮挡,实现跨不同光照和遮挡模式的泛化。

Comments Project Page: https://genwildsplat.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27017 2026-05-01 eess.IV cs.LG stat.ML 71%

Validating the Clinical Utility of CineECG 3D Reconstructions through Cross-Modal Feature Attribution

通过跨模态特征归因验证CineECG 3D重建的临床价值

Karol Dobiczek, Maciej Mozolewski, Szymon Bobek, Michał Szafarczyk, Peter van Dam, Grzegorz J. Nalepa

机构 * Department of Human-Centered Artificial Intelligence, Institute of Applied Computer Science, Jagiellonian University(人类中心人工智能系,应用计算机科学研究所,雅盖隆大学) Faculty of Medicine, Jagiellonian University Medical College(医学系,雅盖隆大学医学院)

专题命中 三维重建 :3D reconstruction(title)

AI总结 本文提出跨模态方法,将高性能12导联ECG模型的特征归因投影到CineECG 3D解剖空间,验证了其在临床整合中的有效性。

Comments Accepted to the CompHealth workshop at the 26th International Conference on Computational Science

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10267 2026-05-01 cs.CV 57%

Long-LRM++: Preserving Fine Details in Feed-Forward Wide-Coverage Reconstruction

Long-LRM++:在前馈宽覆盖重建中保留细节

Chen Ziwen, Hao Tan, Peng Wang, Zexiang Xu, Li Fuxin

机构 * Adobe Research(Adobe研究院) Tripo AI(Tripo人工智能) Hillbot(Hillbot公司) Oregon State University(俄勒冈州立大学)

专题命中 三维重建 :Gaussian Splatting(abstract);分类 cs.CV

AI总结 Long-LRM++通过半显式场景表示与轻量解码器,在保持LaCT渲染质量的同时实现实时14FPS的A100 GPU性能,且能扩展至64输入视角,并在ScanNetv2上优于直接从高斯生成深度预测。

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition Findings (CVPRF), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

2. NeRF 1 篇

2604.27702 2026-05-01 cs.CV 90%

RayFormer: Modeling Inter- and Intra-Ray Similarity for NeRF-Based Video Snapshot Compressive Imaging

RayFormer:基于NeRF的视频快照压缩成像中建模跨射线和内射线相似性

Yubo Dong, Danhua Liu, Anqi Li, Zhenyuan Lin

机构 * Changzhi Medical College(长治医学院) Uniwave Artificial Intelligence Technology Co., Ltd.(Uniwave人工智能技术有限公司) Xidian University, School of Artificial Intelligence(西安电子科技大学人工智能学院)

专题命中 NeRF :NeRF(title,title_cn);分类 cs.CV

AI总结 本文提出了一种基于NeRF的视频快照压缩成像方法,通过改进的射线采样策略和RayFormer模型,提升动态场景重建质量。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. Gaussian Splatting 9 篇

2604.19571 2026-05-01 cs.CV 91%

TransSplat: Unbalanced Semantic Transport for Language-Driven 3DGS Editing

TransSplat:不平衡语义传输用于语言驱动的3DGS编辑

Yanhui Chen, Jiahong Li, Jingchao Wang, Junyi Lin, Zixin Zeng, Yang Shi

机构 * Guangdong University of Technology(广东工业大学) Peking University(北京大学)

专题命中 Gaussian Splatting :3DGS(title,title_cn);Gaussian Splatting(abstract);分类 cs.CV

AI总结 TransSplat通过不平衡语义传输解决语言驱动3DGS编辑中2D证据与3D高斯之间的语义对应问题,提升局部编辑精度和结构一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27552 2026-05-01 cs.CV 89%

Residual Gaussian Splatting for Ultra Sparse-View CBCT Reconstruction

残差高斯点云法用于超稀疏视角CBCT重建

Jian Lin, Jiancheng Fang, Shaoyu Wang, Changan Lai, Yikun Zhang, Yang Chen, Qiegen Liu

机构 * School of Information Engineering, Nanchang University(南昌大学信息工程学院)

专题命中 Gaussian Splatting :3DGS(summary_cn,abstract);Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 本文提出残差高斯点云法,结合小波多分辨率分析与3DGS,解决超稀疏视角CBCT重建中光谱偏倚问题,提升图像细节与几何纹理的精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10223 2026-05-01 cs.AR cs.GR eess.IV 85%

A 129FPS Full HD Real-Time Accelerator for 3D Gaussian Splatting

为3D高斯散射实现129fps全高清实时加速器

Fang-Chi Chang, Tian-Sheuan Chang

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract,abstract_cn);分类 cs.GR

AI总结 本文提出一种低功耗低成本的3D高斯散射硬件加速器,结合迭代高斯剪枝与微调、渐进球谐函数度缩减和所有球谐函数系数及颜色的向量量化,实现全高清实时渲染,并在面积、吞吐量和能效方面优于现有方案。

Journal ref IEEE Transactions on Visualization and Computer Graphics, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.28016 2026-05-01 cs.CV cs.GR cs.LG 81%

Faster 3D Gaussian Splatting Convergence via Structure-Aware Densification

通过结构感知密集化实现更快的3D高斯散射收敛

Linjie Lyu, Ayush Tewari, Jianchun Chen, Thomas Leimkühler, Christian Theobalt

机构 * Cambridge University(剑桥大学) Saarbrücken Research Center for Visual Computing, Interaction, and Artificial Intelligence (VIA)(萨尔布吕肯视觉计算、交互与人工智能研究中心(VIA))

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV、cs.GR

AI总结 本文提出一种结构感知密集化框架,通过多尺度频率分析和各向异性分裂提升3D高斯散射的收敛速度和重建质量。

Comments Siggraph 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.28179 2026-05-01 cs.CV 79%

Stop Holding Your Breath: CT-Informed Gaussian Splatting for Dynamic Bronchoscopy

停止屏息:基于CT的高斯点云法用于动态支气管镜

Andrea Dunn Beltran, Daniel Rho, Aarav Mehta, Xinqi Xiong, Raúl San José Estépar, Ron Alterovitz, Marc Niethammer, Roni Sengupta

机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) Harvard Medical School(哈佛医学院) University of California, San Diego(加州大学圣地亚哥分校)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 本文提出利用患者特定的呼吸模型消除屏息协议需求,通过配对呼气吸气CT扫描减少呼吸运动,实现连续的变形感知重建。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27572 2026-05-01 cs.GR 74%

SandSim: Curve-Guided Gaussian Splatting for Reconstructing Sand Painting Processes

SandSim:基于曲线引导的高斯点云法重建沙画创作过程

Yilin Wang, Haojie Huang, Chen Li, Yang Li, Changbo Wang, Chenhui Li

专题命中 Gaussian Splatting :Gaussian Splatting(title);分类 cs.GR

AI总结 本文提出SandSim框架,通过曲线引导的高斯表示和减法合成方案,实现单张图像中沙画创作过程的重建,提升时空一致性和视觉真实感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27590 2026-05-01 cs.CV 70%

Fake3DGS: A Benchmark for 3D Manipulation Detection in Neural Rendering

Fake3DGS:3D操控检测在神经渲染中的基准

Davide Di Nucci, Riccardo Catalini, Guido Borghi, Roberto Vezzani

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出Fake3DGS基准,用于评估3D内容的真实性,通过3D高斯点散布场景和渲染视图,展示现有2D检测器在区分原始与3D操控图像上的不足,并引入3D-aware方法提升识别性能。

Comments Accepted at ICPR 2026. Code and data: https://github.com/iot-unimore/Fake3DGS

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.27437 2026-05-01 cs.CV 70%

Softmax-GS: Generalized Gaussians Learning When to Blend or Bound

Softmax-GS: 通用高斯学习何时融合或边界

Chen Ziwen, Peng Wang, Hao Tan, Zexiang Xu, Li Fuxin

机构 * Adobe Research(Adobe研究院) Tripo AI(Tripo人工智能公司) Hillbot(Hillbot公司) Oregon State University(俄勒冈州立大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 Softmax-GS通过在重叠区域引入softmax竞争机制,解决3D GS中的视图不一致和模糊边界问题,提升重建质量和参数效率。

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition Findings (CVPRF), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.28122 2026-05-01 cs.CV cs.LG 57%

Beyond Gaussian Bottlenecks: Topologically Aligned Encoding of Vision-Transformer Feature Spaces

超越高斯瓶颈:基于拓扑对齐的视觉Transformer特征空间编码

Andrew Bond, Ilkin Umut Melanlioglu, Erkut Erdem, Aykut Erdem

机构 * Department of Computer Engineering, Koç University, Istanbul, Turkey(科克大学计算机工程系,伊斯坦布尔,土耳其) Department of Computer Engineering, Hacettepe University, Ankara, Turkey(哈恰塔佩大学计算机工程系,安卡拉,土耳其) KUIS AI Research Center, Istanbul, Turkey(KUIS人工智能研究中心,伊斯坦布尔,土耳其) Department of Electrical and Electronics Engineering, Koç University, Istanbul, Turkey(科克大学电气与电子工程系,伊斯坦布尔,土耳其)

专题命中 Gaussian Splatting :point cloud(abstract);分类 cs.CV

AI总结 本文提出S²VAE框架,通过压缩和表示场景的3D状态,包括相机运动、深度和点结构,以提升视觉模型的几何一致性。实验显示,几何对齐的超球面隐空间在高压缩条件下优于传统高斯瓶颈。

Comments 16 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 点云 3 篇

2602.00937 2026-05-01 cs.RO cs.AI cs.CV cs.LG 62%

CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining

CLAMP: 基于对比学习的3D多视角动作条件机器人操控预训练

I-Chun Arthur Liu, Krzysztof Choromanski, Sandy Huang, Connor Schenck

机构 * Google DeepMind(谷歌DeepMind) University of Southern California(南加州大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 CLAMP通过3D点云和机器人动作进行预训练,利用对比学习提升机器人操控精度与效率,优于现有基线方法。

Comments Accepted to the Robotics: Science and Systems (RSS) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13486 2026-05-01 cs.CV 57%

Uncertainty Quantification Framework for Aerial and UAV Photogrammetry through Error Propagation

通过误差传播的航空和无人机摄影测量不确定性量化框架

Debao Huang, Rongjun Qin

机构 * Geospatial Data Analytics Laboratory, The Ohio State University(地理空间数据分析实验室,俄亥俄州立大学) Department of Civil, Environmental and Geodetic Engineering, The Ohio State University(土木、环境与大地测量工程系,俄亥俄州立大学) Department of Electrical and Computer Engineering, The Ohio State University(电气与计算机工程系,俄亥俄州立大学) Translational Data Analytics Institute, The Ohio State University(转化数据分析研究院,俄亥俄州立大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出通过误差传播的不确定性量化框架,解决多视立体阶段的不确定性估计问题,利用自校准方法提升摄影测量点云的鲁棒性和可验证性。

Comments 27 pages, 12 figures, this manuscript has been accepted to ISPRS Journal of Photogrammetry and Remote Sensing

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.28040 2026-05-01 eess.SP 50%

LiDAR-based Dynamic Blockage Prediction: A Data-driven Approach for Learning Interactive Bayesian Models

基于LiDAR的动态遮挡预测:一种数据驱动的方法用于学习交互式贝叶斯模型

Saleemullah Memon, Ali Krayani, Pamela Zontone, Lucio Marcenaro, David Martin Gomez, Carlo Regazzoni

专题命中 点云 :point cloud(abstract)

AI总结 本文提出了一种数据驱动方法,用于学习交互式通用动态贝叶斯网络模型,以从基于时间序列的3D点云感知预测未来LiDAR传感器遮挡。通过训练不同车辆在正常和遮挡情况下的GDBN模型,并利用高阶词汇进行多车辆交互,提出交互式马尔可夫跳跃粒子滤波器以推断遮挡并检测异常。

Comments 2025 IEEE International Workshop on Technologies for Defense and Security (TechDefense), Rome, Italy

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 空间理解 1 篇

2604.28196 2026-05-01 cs.CV 57%

HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation

HERMES++:迈向统一的驾驶世界模型用于3D场景理解和生成

Xin Zhou, Dingkang Liang, Xiwu Chen, Feiyang Tan, Dingyuan Zhang, Hengshuang Zhao, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学) Mach Drive University of Hong Kong(香港大学)

专题命中 空间理解 :point cloud(abstract);分类 cs.CV

AI总结 本文提出HERMES++,一种统一的驾驶世界模型,整合3D场景理解和未来几何预测。通过BEV表示、LLM增强世界查询和当前到未来链接等设计,提升驾驶场景的生成与理解能力。

Comments Extended version of ICCV 25 paper HERMES, Code: https://github.com/H-EmbodVis/HERMESV2, Project page: https://h-embodvis.github.io/HERMESV2/

详情

展开后加载摘要…

URL PDF HTML 收藏