arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-03-09 至 2026-03-09 共收录 27 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 6 篇

2601.01144 2026-03-09 cs.RO 83%

VISO: Robust Underwater Visual-Inertial-Sonar SLAM with Photometric Rendering for Dense 3D Reconstruction

VISO:具有光度渲染的鲁棒水下视觉-惯性-声纳SLAM用于密集3D重建

Shu Pan, Simon Archieri, Ahmet Cinar, Jonatan Scharff Willners, Ignacio Carlucho, Yvan Petillot

机构 * School of Engineering and Physical Sciences, Heriot-Watt University(工程与物理科学学院,赫瑞斯泰大学) Frontier Robotics, The National Robotarium(前沿机器人,国家机器人arium)

专题命中 三维重建 :3D reconstruction(title,abstract);point cloud(abstract);分类 cs.RO

AI总结 VISO通过融合视觉、惯性与声纳数据,结合光度渲染技术,实现高精度水下6自由度定位和实时密集3D重建。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06374 2026-03-09 cs.CV 70%

Rewis3d: Reconstruction Improves Weakly-Supervised Semantic Segmentation

Rewis3d:重建提升弱监督语义分割

Jonas Ernst, Wolfgang Boettcher, Lukas Hoyer, Jan Eric Lenssen, Bernt Schiele

机构 * Saarland University(萨尔兰大学) Max Planck Institute for Informatics(马克斯·普朗克信息研究所) ETH Zurich(苏黎世联邦理工学院)

专题命中 三维重建 :3D reconstruction(abstract);point cloud(abstract);分类 cs.CV

AI总结 Rewis3d通过利用3D重建作为辅助监督信号,提升弱监督语义分割的性能,优于现有方法2-7%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18923 2026-03-09 cs.CV 70%

Gaussian Set Surface Reconstruction through Per-Gaussian Optimization

通过每高斯优化进行高斯集表面重建

Zhentao Huang, Di Wu, Zhenbang He, Minglun Gong

机构 * School of Computer Science, University of Guelph(圭尔夫大学计算机科学学院) Faculty of Science and Technology, University of Macau(澳门大学科技学院) Irving K. Barber Faculty of Science, UBC Okanagan(UBC Okanagan科学学院)

专题命中 三维重建 :Gaussian Splatting(abstract);3DGS(abstract);分类 cs.CV

AI总结 GSSR通过每高斯优化实现高斯点均匀分布与法线对齐,提升3D场景重建精度和编辑效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05604 2026-03-09 cs.CV cs.LG cs.RO 62%

From Decoupled to Coupled: Robustness Verification for Learning-based Keypoint Detection with Joint Specifications

从解耦到耦合:基于联合规范的学习关键点检测的鲁棒性验证

Xusheng Luo, Changliu Liu

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV、cs.RO

AI总结 本文提出首个基于联合规范的学习关键点检测鲁棒性验证框架,通过混合整数线性规划实现对关键点联合偏差的约束,提升验证效率和鲁棒性保证。

Comments 21 pages, 4 figures, 9 tables. arXiv admin note: text overlap with arXiv:2408.00117

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05787 2026-03-09 cs.CV 57%

Spectral Probing of Feature Upsamplers in 2D-to-3D Scene Reconstruction

特征上采样器在2D到3D场景重建中的频谱探测

Ling Xiao, Yuliang Xiu, Yue Chen, Guoming Wang, Toshihiko Yamasaki

机构 * Hokkaido University(北海道大学) Westlake University(西湖大学) Zhejiang University(浙江大学) The University of Tokyo(东京大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出频谱诊断框架,通过六个指标评估特征上采样器对3D重建的影响,发现频谱一致性比空间细节更关键。

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.15048 2026-03-09 cs.CV cs.AI cs.LG cs.MM 57%

Make VLM Recognize Visual Hallucination on Cartoon Character Image with Pose Information

使VLM在卡通角色图像上通过姿态信息识别视觉幻觉

Bumsoo Kim, Wonseop Shin, Kyuchul Lee, Yonghoon Jung, Sanghyun Seo

机构 * Chung-Ang University(Chung-Ang 大学) Coupang(韩国Coupang)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出基于姿态信息的VLM幻觉检测方法,通过上下文学习提升识别准确率,实验表明在卡通角色图像中幻觉检测效果提升50%-80%。

Comments Accepted at WACV 2025, Project page: https://gh-bumsookim.github.io/Cartoon-Hallucinations-Detection/. (Fixed typos)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. NeRF 1 篇

2601.03869 2026-03-09 cs.CV cs.GR cs.LG cs.RO 67%

Bayesian Monocular Depth Refinement via Neural Radiance Fields

基于神经辐射场的贝叶斯单目深度细化

Arun Muthukkumar

机构 * Department of Computer Science Illinois Mathematics Science Academy Aurora, United States

专题命中 NeRF :NeRF(abstract);分类 cs.CV、cs.GR、cs.RO

AI总结 本文提出MDENeRF,通过神经辐射场和贝叶斯融合改进单目深度估计,提升场景理解的精细几何细节。

Comments IEEE 8th International Conference on Algorithms, Computing and Artificial Intelligence (ACAI 2025)

Journal ref Proc. IEEE 8th International Conference on Algorithms, Computing and Artificial Intelligence (ACAI), pp. 488-492, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

3. Gaussian Splatting 7 篇

2603.05882 2026-03-09 cs.CV 89%

CylinderSplat: 3D Gaussian Splatting with Cylindrical Triplanes for Panoramic Novel View Synthesis

CylinderSplat:基于圆柱三平面的3D高斯散射用于全景新视角合成

Qiwei Wang, Xianghui Ze, Jingyi Yu, Yujiao Shi

机构 * Shanghaitech University(上海科技大学) Nanjing University of Science and Technology(南京理工大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);novel view synthesis(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 CylinderSplat通过圆柱三平面表示和双分支架构,提升全景图像新视角合成的重建质量和几何精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06216 2026-03-09 cs.CV 85%

EntON: Eigenentropy-Optimized Neighborhood Densification in 3D Gaussian Splatting

EntON:在3D高斯点云中基于特征熵的邻域稠密化

Miriam Jäger, Boris Jutzi

机构 * Institute of Photogrammetry and Remote Sensing (IPF)(摄影测量与遥感研究所(IPF)) Karlsruhe Institute of Technology (KIT)(卡尔斯鲁厄理工学院)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 EntON通过Eigenentropy优化邻域稠密化,提升3D重建的几何精度和渲染质量,同时减少高斯数量和训练时间。

Comments Submitted to ISPRS Journal of Photogrammetry and Remote Sensing on 20 February 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.17655 2026-03-09 cs.CV 85%

FeatureGS: Eigenvalue-Feature Optimization in 3D Gaussian Splatting for Geometrically Accurate and Artifact-Reduced Reconstruction

FeatureGS: 3D高斯散射中基于特征值的特征优化以实现几何准确且减少伪影的重建

Miriam Jäger, Markus Hillemann, Boris Jutzi

机构 * Institute of Photogrammetry and Remote Sensing(摄影测量与遥感研究所) Karlsruhe Institute of Technology(卡尔斯鲁厄理工大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);point cloud(abstract);分类 cs.CV

AI总结 FeatureGS通过引入基于特征值的几何损失项,提高3DGS的几何准确性,减少伪影和内存需求,实现更高效的3D场景重建。

Comments 16 pages, 9 figures, 7 tables

Journal ref ISPRS Open Journal of Photogrammetry and Remote Sensing Volume 17, August 2025, 100100

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06061 2026-03-09 cs.CV cs.RO 84%

Transforming Omnidirectional RGB-LiDAR data into 3D Gaussian Splatting

将全方位RGB-LiDAR数据转换为3D高斯散点

Semin Bae, Hansol Lim, Jongseong Brad Choi

机构 * Department of Computer Science, State University of New York(计算机科学系,纽约州立大学) Department of Mechanical Engineering, State University of New York(机械工程系,纽约州立大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV、cs.RO

AI总结 本文提出了一种将全方位RGB-LiDAR数据转换为3D高斯散点的重用管道,解决数据处理中的非线性失真和计算开销问题,提升复杂场景的渲染保真度。

Comments This work has been submitted to the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05932 2026-03-09 cs.CV cs.RO 82%

FTSplat: Feed-forward Triangle Splatting Network

FTSplat:前馈三角形散射网络

Xiong Jinlin, Li Can, Shen Jiawei, Qi Zhigang, Sun Lei, Zhao Dongyang

专题命中 Gaussian Splatting :NeRF(abstract);Gaussian Splatting(abstract);3DGS(abstract);point cloud(abstract)

AI总结 FTSplat通过前馈三角形生成网络实现高效的三维重建,无需每场景优化,直接生成可用于模拟的连续三角形表面。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04843 2026-03-09 cs.CV 77%

PoI: A Filter to Extract Pixel of Interest from Novel Views for Scene Coordinate Regression

PoI: 一种从新视角提取感兴趣像素的滤波器用于场景坐标回归

Feifei Li, Qi Song, Chi Zhang, Hui Shuai, Rui Huang

专题命中 Gaussian Splatting :NeRF(abstract);Gaussian Splatting(abstract);3DGS(abstract);分类 cs.CV

AI总结 PoI通过像素级过滤策略提升场景坐标回归的定位精度,结合扩散模型和重投影误差优化新视角合成效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11689 2026-03-09 cs.RO cs.AI 57%

Phys2Real: Fusing VLM Priors with Interactive Online Adaptation for Uncertainty-Aware Sim-to-Real Manipulation

Phys2Real: 融合视觉语言模型先验与交互在线适应以实现不确定性感知的仿真到现实操控

Maggie Wang, Stephen Tian, Aiden Swann, Ola Shorinwa, Jiajun Wu, Mac Schwager

机构 * Stanford University(斯坦福大学) Princeton University(普林斯顿大学)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.RO

AI总结 Phys2Real通过融合视觉语言模型先验与交互适应,提升仿真到现实操控的不确定性感知与任务成功率。

Comments Accepted to IEEE International Conference on Robotics and Automation (ICRA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 点云 9 篇

2603.06321 2026-03-09 cs.CV 79%

P-SLCR: Unsupervised Point Cloud Semantic Segmentation via Prototypes Structure Learning and Consistent Reasoning

P-SLCR:通过原型结构学习和一致推理实现无监督点云语义分割

Lixin Zhan, Jie Jiang, Tianjian Zhou, Yukun Du, Yan Zheng, Xuehu Duan

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 P-SLCR通过原型结构学习和一致推理,实现点云语义分割的无监督方法,取得优于全监督方法的性能。

Journal ref AAAI2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05858 2026-03-09 cs.CR 71%

Indoor Space Authentication by ISS-based Keypoint Extraction from 3D Point Clouds

基于3D点云的内在形状特征提取的室内空间认证

Yuki Yamada, Daisuke Kotani, Kota Tsubouchi, Hidehito Gomi, Yasuo Okabe

专题命中 点云 :point cloud(title)

AI总结 ISS-RegAuth通过稀疏关键点提取实现轻量级室内空间认证,降低误差率和数据传输量,提升隐私保护与边缘部署能力。

Comments Accepted in IEEE PerCom 2026 as a Work-in-Progress paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16538 2026-03-09 cs.CV 70%

OnlineSI: Taming Large Language Model for Online 3D Understanding and Grounding

OnlineSI: 通过大规模语言模型实现在线3D理解和定位

Zixian Liu, Zhaoxi Chen, Liang Pan, Ziwei Liu

机构 * Tsinghua University(清华大学) Nanyang Technological University(南洋理工大学) Shanghai AI Lab(上海人工智能实验室)

专题命中 点云 :point cloud(abstract);spatial understanding(abstract);分类 cs.CV

AI总结 OnlineSI通过整合3D点云与语义信息,提升大规模语言模型在动态环境中的空间理解和物体识别能力。

Comments Project Page: https://onlinesi.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06512 2026-03-09 cs.RO cs.CV 62%

SG-DOR: Learning Scene Graphs with Direction-Conditioned Occlusion Reasoning for Pepper Plants

SG-DOR:基于方向条件遮挡推理的学习场景图用于Pepper植物

Rohit Menon, Niklas Mueller-Goldingen, Sicong Pan, Gokul Krishna Chenchani, Maren Bennewitz

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 SG-DOR通过方向条件遮挡推理,提升密集作物中果实遮挡预测与连接推断的准确性,为机器人采摘提供结构化关系信号。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06250 2026-03-09 cs.CV 57%

Hierarchical Collaborative Fusion for 3D Instance-aware Referring Expression Segmentation

层次化协作融合用于3D实例感知指代表达分割

Keshen Zhou, Runnan Chen, Mingming Gong, Tongliang Liu

机构 * The University of Sydney(悉尼大学) The University of Melbourne(墨尔本大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 HCF-RES通过层次化视觉语义分解和渐进多级融合,实现了3D实例感知指代表达分割的高精度与细粒度定位。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05355 2026-03-09 cs.RO 57%

OmniDP: Beyond-FOV Large-Workspace Humanoid Manipulation with Omnidirectional 3D Perception

OmniDP: 在超广角大工作空间中实现人形机器人的全方位3D感知

Pei Qu, Zheng Li, Yufei Jia, Ziyun Liu, Liang Zhu, Haoang Li, Jinni Zhou, Jun Ma

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Tsinghua University(清华大学)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 OmniDP通过360度全景点云感知和时间感知注意力池化机制,实现人形机器人在大工作空间中的稳健操作,优于传统深度相机方法。

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01034 2026-03-09 cs.CV cs.AI cs.LG 57%

Reparameterized Tensor Ring Functional Decomposition for Multi-Dimensional Data Recovery

重新参数化张量环功能分解用于多维数据恢复

Yangyang Xu, Junbo Ke, You-Wei Wen, Chao Wang

机构 * Key Laboratory of Computing and Stochastic Mathematics (Ministry of Education)(计算与随机数学重点实验室(教育部)) School of Mathematics and Statistics(数学与统计学学院) Department of Statistics and Data Science(统计与数据科学系)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出了一种重新参数化的张量环功能分解方法,通过结合可学习的潜在张量和固定基底,提升多维数据恢复的性能。

Comments 22 pages, 18 figures, 12 tables. Accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06496 2026-03-09 physics.med-ph 50%

Rotation-invariant graph message passing enables acquisition protocol generalisation in learning-based brain microstructure estimation

旋转不变图消息传递实现基于学习的脑微结构估计中的获取协议泛化

Leevi Kerkelä, Hui Zhang

专题命中 点云 :point cloud(abstract)

AI总结 本文提出一种旋转不变图消息传递网络,通过模拟数据训练实现微结构估计的协议泛化,无需重新训练即可适应未知协议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06380 2026-03-09 math.NA cs.AI cs.LG cs.NA 50%

Kinetic-based regularization: Learning spatial derivatives and PDE applications

基于动力学的正则化:学习空间导数和PDE应用

Abhisek Ganguly, Santosh Ansumali, Sauro Succi

机构 * Engineering Mechanics Unit(工程力学单元) Jawaharlal Nehru Centre for Advanced Scientific Research(贾瓦哈拉尔·奈尔中心先进科学研究所) Fondazione Istituto Italiano di Tecnologia(意大利技术研究院)

专题命中 点云 :point cloud(abstract)

AI总结 本文提出基于动力学的正则化方法,用于高精度学习空间导数并应用于偏微分方程的求解,通过两种方案实现二次收敛性,提升在不规则点云上求解PDE的稳定性与准确性。

Comments Published as a conference paper at ICLR 2026 Workshop AI and PDE

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 3D生成 2 篇

2603.05845 2026-03-09 cs.CV 79%

Cog2Gen3D: Sculpturing 3D Semantic-Geometric Cognition for 3D Generation

Cog2Gen3D: 三维语义-几何认知雕刻用于三维生成

Haonan Wang, Hanyu Zhou, Haoyue Liu, Tao Gu, Luxin Yan

机构 * School of Artificial and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院) School of Computing, National University of Singapore(新加坡国立大学计算机学院) School of Computing, Macquarie University(麦考瑞大学计算机学院)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 Cog2Gen3D通过结合语义和绝对几何信息,提出了一种三维认知引导的扩散框架,以实现可控的三维生成。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08535 2026-03-09 cs.CV 79%

Photo3D: Advancing Photorealistic 3D Generation through Structure-Aligned Detail Enhancement

Photo3D: 通过结构对齐的细节增强推进光实3D生成

Xinyue Liang, Zhinyuan Ma, Lingchen Sun, Yanjun Guo, Lei Zhang

机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 Photo3D通过结构对齐的细节增强技术,提升3D生成的逼真度,适用于多种3D原生生成范式。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. SLAM与定位 2 篇

2601.05805 2026-03-09 cs.RO 57%

InsSo3D: Inertial Navigation System and 3D Sonar SLAM for turbid environment inspection

InsSo3D:惯性导航系统与3D声呐SLAM用于浑浊环境检测

Simon Archieri, Ahmet Cinar, Shu Pan, Jonatan Scharff Willners, Michele Grimaldi, Ignacio Carlucho, Yvan Petillot

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.RO

AI总结 InsSo3D结合3D声呐和惯性导航系统,实现高效准确的水下SLAM,适用于浑浊环境中的结构检测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06306 2026-03-09 cs.CV cs.AI 57%

Exploiting Spatiotemporal Properties for Efficient Event-Driven Human Pose Estimation

利用时空特性实现高效事件驱动的人体姿态估计

Haoxian Zhou, Chuanzhi Xu, Langyi Chen, Pengfei Ye, Haodong Chen, Yuk Ying Chung, Qiang Qu

机构 * The University of Sydney(悉尼大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 SLAM与定位 :point cloud(abstract);分类 cs.CV

AI总结 本文提出利用事件流的时空特性,结合点云框架提升人体姿态估计性能,实验显示在DHP19数据集上平均MPJPE减少4%。

详情

展开后加载摘要…

URL PDF HTML 收藏