arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-04-30 至 2026-04-30 共收录 20 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 3 篇

2604.04135 2026-04-30 cs.CV 57%

NTIRE 2026 3D Restoration and Reconstruction in Real-world Adverse Conditions: RealX3D Challenge Results

NTIRE 2026 3D修复与重建在真实恶劣环境中的应用:RealX3D挑战赛结果

Shuhong Liu, Chenyu Bao, Ziteng Cui, Xuangeng Chu, Bin Ren, Lin Gu, Xiang Chen, Mingrui Li, Long Ma, Marcos V. Conde, Radu Timofte, Yun Liu, Ryo Umagami, Tomohiro Hashimoto, Zijian Hu, Yuan Gan, Tianhan Xu, Yusuke Kurose, Tatsuya Harada, Junwei Yuan, Gengjia Chang, Xining Ge, Mache You, Qida Cao, Zeliang Li, Xinyuan Hu, Hongde Gu, Changyue Shi, Jiajun Ding, Zhou Yu, Jun Yu, Seungsang Oh, Fei Wang, Donggun Kim, Zhiliang Wu, Seho Ahn, Xinye Zheng, Kun Li, Yanyan Wei, Weisi Lin, Dizhe Zhang, Yuchao Chen, Meixi Song, Hanqing Wang, Haoran Feng, Lu Qi, Jiaao Shan, Yang Gu, Jiacheng Liu, Shiyu Liu, Kui Jiang, Junjun Jiang, Runyu Zhu, Sixun Dong, Qingxia Ye, Zhiqiang Zhang, Zhihua Xu, Zhiwei Wang, Phan The Son, Zhimiao Shi, Zixuan Guo, Xueming Fu, Lixia Han, Changhe Liu, Zhenyu Zhao, Manabu Tsukada, Zheng Zhang, Zihan Zhai, Tingting Li, Ziyang Zheng, Yuhao Liu, Dingju Wang, Jeongbin You, Younghyuk Kim, Il-Youp Kwak, Mingzhe Lyu, Junbo Yang, Wenhan Yang, Hongsen Zhang, Jinqiang Cui, Hong Zhang, Haojie Guo, Hantang Li, Qiang Zhu, Bowen He, Xiandong Meng, Debin Zhao, Xiaopeng Fan, Wei Zhou, Linzhe Jiang, Linfeng Li, Louzhe Xu, Qi Xu, Hang Song, Chenkun Guo, Weizhi Nie, Yufei Li, Xingan Zhan, Zhanqi Shi, Dufeng Zhang, Boyuan Tian, Jingshuo Zeng, Gang He, Yubao Fu, Weijie Wang, Cunchuan Huang

机构 * XInsight Lab(XInsight实验室) Hunan Duo(湖南 Duo) Insta3D DLMath_Vision(DLMath Vision) Insta3DDD Wanderer Diouj. El K7MQ Windrise ZZZ LowLight Wizards SUSTech-PCL(四川大学-PCL) EE-GS DV-Lowlight(3DV-Lowlight) DSmokeR(3DSmokeR) HangFans MonoSmokeGS FJNU-STAR DLLR(3DLLR) HNU AAA Harmony3D RunAI AIC-GER CC

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文综述了NTIRE 2026 3D修复与重建挑战赛,评估了提交方法在极端低光和烟雾降级环境下的3D重建进展,揭示了顶级方法的共同设计原则和有效策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26232 2026-04-30 cs.CV cs.AI 57%

DepthPilot: From Controllability to Interpretability in Colonoscopy Video Generation

DepthPilot:从可控性到可解释性在结肠镜视频生成中

Junhu Fu, Ke Chen, Weidong Guo, Shuyu Liang, Jie Xu, Chen Ma, Kehao Wang, Shengli Lin, Zeju Li, Yuanyuan Wang, Yi Guo, Shuo Li

机构 * College of Biomedical Engineering, Fudan University, Shanghai 200433, China(复旦大学生物医学工程学院) Key Laboratory of Medical Imaging Computing and Computer Assisted Intervention of Shanghai, Shanghai 200032, China(上海医学影像计算与计算机辅助干预重点实验室) Endoscopy Research Institute, Zhongshan Hospital, Fudan University, Shanghai 200032, China(复旦大学中山医院内窥镜研究所) Shanghai Collaborative Innovation Center of Endoscopy, Shanghai 200032, China(上海内窥镜协同创新中心) Department of Biomedical Engineering, Case Western Reserve University, Cleveland, OH 44106, USA(凯斯西储大学生物医学工程系) Department of Computer and Data Science, Case Western Reserve University, Cleveland, OH 44106, USA(凯斯西储大学计算机与数据科学系)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出DepthPilot框架,通过几何对齐策略和自适应样条去噪模块,实现结肠镜视频生成的可控性与可解释性,取得高FID分数和临床评估领先成果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22658 2026-04-30 cs.CV 57%

PASR: Pose-Aware 3D Shape Retrieval from Occluded Single Views

PASR:基于遮挡单视图的姿态感知3D形状检索

Jiaxin Shi, Guofeng Zhang, Wufei Ma, Naifu Liang, Adam Kortylewski, Alan Yuille

机构 * Shanghai Jiao Tong University(上海交通大学) Johns Hopkins University(约翰霍普金斯大学) University of California, San Diego(加州大学圣地亚哥分校) CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全中心)

专题命中 三维重建 :point cloud(abstract);分类 cs.CV

AI总结 PASR通过将2D基础模型知识蒸馏到3D编码器,解决单视图3D形状检索中姿态感知与遮挡问题,实现更鲁棒的检索与多任务能力。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. NeRF 1 篇

2604.26899 2026-04-30 eess.SY cs.RO cs.SY 57%

Safe Navigation using Neural Radiance Fields via Reachable Sets

通过可达集实现神经辐射场的安全导航

Omanshu Thapliyal, Malarvizhi Sankaranarayanasamy, Ravigopal Vennelakanti

机构 * Researcher, Hitachi America Ltd.(Hitachi America Ltd.研究人员) Sr. Researcher, Hitachi America Ltd.(Hitachi America Ltd.高级研究人员)

专题命中 NeRF :NeRF(abstract_cn);分类 cs.RO

AI总结 本文提出利用可达集和神经辐射场进行安全导航,通过约束最优控制解决路径规划问题,验证了在复杂环境中安全避障的能力。

Comments 5 pages, 8 figures, 2026 4th International Conference on Mechatronics, Control and Robotics (ICMCR)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. Gaussian Splatting 5 篇

2604.16747 2026-04-30 cs.CV 89%

Incoherent Deformation, Not Capacity: Diagnosing and Mitigating Overfitting in Dynamic Gaussian Splatting

不一致变形,而非容量:诊断和缓解动态高斯点云中的过拟合

Ahmad Droby

机构 * Independent Researcher(独立研究者)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);NeRF(abstract,abstract_cn);3DGS(abstract,abstract_cn);分类 cs.CV

AI总结 本文发现动态高斯点云过拟合源于不一致变形而非参数量,通过引入弹性能量正则化等方法显著减少过拟合差距。

Comments 10 pages, 6 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26238 2026-04-30 cs.CV 85%

EnerGS: Energy-Based Gaussian Splatting with Partial Geometric Priors

EnerGS: 基于部分几何先验的能量场能量基于高斯点散射

Rui Song, Tianhui Cai, Markus Gross, Yun Zhang, Walter Zimmer, Zhiyu Huang, Olaf Wysocki, Jiaqi Ma

机构 * University of California, Los Angeles, California, USA(加州大学洛杉矶分校) University of Cambridge, Cambridge, UK(剑桥大学) Technical University of Munich, Munich, Germany(慕尼黑技术大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract,abstract_cn);分类 cs.CV

AI总结 EnerGS通过建模部分可观测几何为连续能量场,提升大规模户外场景中高斯点散射的光度质量和几何稳定性,同时缓解过拟合问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25945 2026-04-30 eess.SP cs.AI 85%

Planar Gaussian Splatting with Bilinear Spatial Transformer for Wireless Radiance Field Reconstruction

平面高斯散射与双线性空间变换器用于无线辐射场重建

Jinghan Zhang, Xitao Gong, Qi Wang, Richard A. Stirling-Gallacher, Giuseppe Caire

机构 * Munich Research Center, Huawei Technologies Duesseldorf GmbH, Munich, Germany Department of Electrical Engineering(慕尼黑研究中心,华为技术杜塞尔多夫有限公司,慕尼黑,德国电气工程系) Computer Science, Technical University of Berlin, Berlin, Germany(计算机科学,柏林技术大学,柏林,德国)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);NeRF(abstract,abstract_cn)

AI总结 本文提出BiSplat-WRF,通过去除冗余投影并引入全局电磁耦合,提升无线辐射场重建的物理可解释性和精度,实验显示其在结构相似性指数上优于现有方法。

Comments Accepted to IEEE ICC 2026 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.19216 2026-04-30 cs.NI cs.AI cs.CV cs.LG 77%

Bridging Visual and Wireless Sensing via a Unified Radiation Field for 3D Radio Map Construction

通过统一的辐射场桥接视觉与无线传感以实现3D无线电地图构建

Chaozheng Wen, Jingwen Tong, Zehong Lin, Chenghong Bian, Jun Zhang

机构 * Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 Gaussian Splatting :NeRF(abstract,abstract_cn);Gaussian Splatting(abstract);分类 cs.CV

AI总结 本文提出URF-GS框架,结合3D高斯散射和反渲染技术,融合多模态数据以提升3D无线电地图的空间频谱精度和样本效率。

Comments The code for this work will be publicly available at: https://github.com/wenchaozheng/URF-GS

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26920 2026-04-30 cs.CV 57%

Color-Encoded Illumination for High-Speed Volumetric Scene Reconstruction

彩色编码照明用于高速体素场景重建

David Novikov, Eilon Vaknin, Narek Tumanyan, Mark Sheinin

机构 * Weizmann Institute of Science(魏茨曼科学研究院)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 本文提出一种无需修改相机硬件的高速体素重建方法,通过快速序列彩色照明实现多视角同时捕获,利用动态高斯点扩散技术解码时空信息,实现实时体素场景重建。

Comments accepted to IEEE CVPR 2026 as a highlight

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 点云 6 篇

2604.26318 2026-04-30 cs.CV 79%

Point Cloud Registration via Probabilistic Self-Update Local Correspondence and Line Vector Sets

点云配准 via 概率自更新局部对应与线向量集

Kuo-Liang Chung, Yu-Cheng Lin, Wu-Chi Chen

机构 * Department of Computer Science and Information Engineering, National Taiwan University of Science and Technology(资讯工程系,国立台湾科技大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出一种快速有效的点云配准算法,结合概率自更新局部对应与线向量集,通过双RANSAC模型提升精度与效率,实现更优的配准性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.24169 2026-04-30 cs.CV 79%

PointTransformerX: Portable and Efficient 3D Point Cloud Processing without Sparse Algorithms

PointTransformerX:无需稀疏算法的便携式和高效3D点云处理

Laurenz Reichardt, Nikolas Ebert, Oliver Wasenmüller

机构 * Mannheim University of Applied Sciences(曼海姆应用科学大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 PointTransformerX采用全PyTorch原生的视觉Transformer架构,无需定制CUDA运算符,通过3D-GS-RoPE实现高效的3D空间关系编码,提升3D点云处理的准确性和效率,同时在ScanNet上达到98.7%的准确率,参数更少,速度更快,内存更小。

Comments This paper has been accepted at IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26868 2026-04-30 cs.CV 57%

Breaking the Rigid Prior: Towards Articulated 3D Anomaly Detection

打破刚性先验:迈向具有铰链或滑动关节的3D异常检测

Jinye Gan, Bozhong Zheng, Xiaohao Xu, Junye Ren, Zixuan Zhang, Na Ni, Yingna Wu

机构 * ShanghaiTech University(上海科技大学) University of Michigan, Ann Arbor(密歇根大学安娜堡分校)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出ArtiAD基准,通过15229个点云数据和六种结构异常类型,解决铰链或滑动关节物体的3D异常检测问题,引入SPA-SDF方法,通过连续姿态条件隐式场实现刚性先验的替代,提升检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22334 2026-04-30 cs.CV 57%

FILTR: Extracting Topological Features from Pretrained 3D Models

FILTR: 从预训练3D模型中提取拓扑特征

Louis Martinez, Maks Ovsjanikov

机构 * LIX, École Polytechnique, IP Paris(巴黎高等师范学院LIX研究所、IP巴黎)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出FILTR框架,通过预训练3D编码器提取点云的拓扑特征,利用transformer解码器生成持续图,首次实现高效的数据驱动持续图提取。

Comments Project page: https://filtr-topology.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26068 2026-04-30 cs.CG 50%

Calibrated Persistent Homology Tests for High-dimensional Collapse Detection

校准的持续同调测试用于高维坍缩检测

Alexander Kalinowski

专题命中 点云 :point cloud(abstract)

AI总结 研究高维点云中坍缩检测,提出基于持续同调的测试统计量,通过校准非坍缩参考模型来评估不同坍缩机制的效力。

Comments Accepted at CG Week 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.06682 2026-04-30 stat.ME stat.CO 50%

Antimodes and Graphical Anomaly Exploration via Adaptive Depth Quantile Functions

反模态与通过自适应深度分位函数的图形异常探索

Gabriel Chandler, Wolfgang Polonik

专题命中 点云 :point cloud(abstract)

AI总结 本文提出了一种新颖的异常检测方法,基于深度分位函数的扩展,在欧几里得和非欧几里得情况下均表现出竞争力,通过可视化分析提升异常检测性能。

Comments 24 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 新视角合成 2 篇

2507.01110 2026-04-30 cs.GR cs.LG 70%

A LoD of Gaussians: Unified Training and Rendering for Ultra-Large Scale Reconstruction with External Memory

高斯点云的层次结构:用于超大规模重建的统一训练与渲染,借助外部内存

Felix Windisch, Thomas Köhler, Lukas Radl, Mattia D'Urso, Michael Steiner, Dieter Schmalstieg, Markus Steinberger

机构 * Graz University of Technology(格拉茨技术大学) University of Stuttgart(斯图加特大学) Huawei Technologies(华为技术)

专题命中 新视角合成 :Gaussian Splatting(abstract);novel view synthesis(abstract);分类 cs.GR

AI总结 本文提出一种无需分块的高斯点云框架,通过外存存储完整场景并动态流式传输相关高斯点,实现超大规模场景的实时训练与渲染,支持多尺度重建与交互式可视化。

Journal ref Proceedings of SIGGRAPH 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00209 2026-04-30 eess.IV cs.AI cs.CV cs.RO 62%

SurgiSR4K: A High-Resolution Endoscopic Video Dataset for Robotic-Assisted Minimally Invasive Procedures

SurgiSR4K:一种用于机器人辅助微创手术的高分辨率内窥镜视频数据集

Fengyi Jiang, Xiaorui Zhang, Lingbo Jin, Ruixing Liang, Yuxin Chen, Adi Chola Venkatesh, Jason Culman, Tiantian Wu, Lirong Shao, Wenqing Sun, Cong Gao, Hallie McNamara, Jingpei Lu, Omid Mohareri

机构 * Intuitive Surgical, Inc.(Intuitive Surgical公司) Johns Hopkins Medicine Neurosurgery(约翰霍普金斯医学神经外科) Johns Hopkins University Electrical and Computer Engineering(约翰霍普金斯大学电气与计算机工程) University of British Columbia Electrical and Computer Engineering(不列颠哥伦比亚大学电气与计算机工程) Wilford & Kate Bailey Small Animal Teaching Hospital(威尔福德与凯蒂·贝利小动物教学医院)

专题命中 新视角合成 :novel view synthesis(abstract);分类 cs.CV、cs.RO

AI总结 本文提出SurgiSR4K数据集,用于提升微创手术的视觉清晰度和计算机辅助导航精度,包含4K分辨率的手术视频,涵盖镜面反射、工具遮挡等挑战性场景,支持超分辨率、烟雾去除等多任务研究。

Journal ref Machine Learning for Biomedical Imaging, Vol. 3, 2025, pp. 875-885

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 3D生成 1 篇

2604.26943 2026-04-30 cs.CV 79%

ProcFunc: Function-Oriented Abstractions for Procedural 3D Generation in Python

ProcFunc:面向Python的基于Blender的程序化3D生成函数抽象

Alexander Raistrick, Karhan Kayan, Jack Nugent, David Yan, Lingjie Mei, Meenal Parakh, Hongyu Wen, Dylan Li, Yiming Zuo, Erich Liang, Jia Deng

机构 * Princeton University(普林斯顿大学)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 ProcFunc通过提供易用的Python函数库,简化了程序化3D生成的创建、组合、分析和执行过程,支持大规模多样化训练数据生成,并减少VLMs在编辑和生成程序化代码时的错误。

详情

展开后加载摘要…

URL PDF HTML 收藏

7. 空间理解 1 篇

2604.26341 2026-04-30 cs.CV 57%

SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness

SpatialFusion:赋予统一图像生成内在三维几何意识

Haiyi Qiu, Kaihang Pan, Jiacheng Li, Juncheng Li, Siliang Tang, Yueting Zhuang

机构 * Zhejiang University(浙江大学) HiThink Research(HiThink研究院)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.CV

AI总结 本文提出SpatialFusion框架,通过引入混合变压器增强MLLM的三维几何建模能力,并利用深度适配器将显式几何约束注入扩散模型,提升空间感知任务性能。

详情

展开后加载摘要…

URL PDF HTML 收藏

8. SLAM与定位 1 篇

2604.26368 2026-04-30 cs.CV 57%

Seamless Indoor-Outdoor Mapping for INGENIOUS First Responders

无缝室内-室外映射用于INGENIOUS第一响应者

Jürgen Wohlfeil, Henry Meißner, Adrian Schischmanow, Thomas Kraft, Dirk Baumbach, Ines Ernst, Dennis Dahlke

机构 * Institute of Optical Systems (OS), German Aerospace Center (DLR)(光学系统研究所(OS),德国航空航天中心(DLR))

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.CV

AI总结 本文提出一种自动化方法,结合自主飞行测绘系统与手持室内定位系统,通过AprilTags实现室内外3D模型无缝融合,实现实时可视化。

详情

展开后加载摘要…

URL PDF HTML 收藏