arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-03-20 至 2026-03-20 共收录 6 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 6 篇

2603.18493 2026-03-20 cs.CV cs.AI cs.LG 79%

FILT3R: Latent State Adaptive Kalman Filter for Streaming 3D Reconstruction

FILT3R:用于流式3D重建的潜在状态自适应卡尔曼滤波器

Seonghyun Jin, Jong Chul Ye

机构 * KAIST AI(韩国科学技术院人工智能研究所)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 FILT3R通过将递归状态更新转化为token空间的随机状态估计,解决了流式3D重建中状态更新规则不稳定的问题,提升了长期稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19231 2026-03-20 cs.CV 74%

MonoArt: Progressive Structural Reasoning for Monocular Articulated 3D Reconstruction

MonoArt:基于渐进结构推理的单目关节3D重建

Haitian Li, Haozhe Xie, Junxiang Xu, Beichen Wen, Fangzhou Hong, Ziwei Liu

机构 * S-Lab, Nanyang Technological University, 637335 Singapore(南洋理工大学S实验室)

专题命中 三维重建 :3D reconstruction(title);分类 cs.CV

AI总结 MonoArt通过渐进结构推理实现单目关节3D重建,无需外部运动模板或多阶段流程,提升重建精度和推理速度,并扩展至机器人操作和关节场景重建。

Comments Project page: https://lihaitian.com/MonoArt

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17781 2026-03-20 cs.CV cs.GR 73%

LiteGE: Lightweight Geodesic Embedding for Efficient Geodesics Computation and Non-Isometric Shape Correspondence

LiteGE:轻量级测地嵌入用于高效测地线计算和非等距形状对应

Yohanes Yudhi Adikusuma, Qixing Huang, Ying He

专题命中 三维重建 :3D vision(abstract);point cloud(abstract);分类 cs.CV、cs.GR

AI总结 LiteGE通过PCA处理信息体素的无符号距离场样本,构建紧凑且类别感知的形状描述符,实现高效测地线计算和非等距形状对应,显著降低内存使用和推理时间。

Journal ref Proceedings of the 40th AAAI Conference on Artificial Intelligence (AAAI-26), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.02691 2026-03-20 cs.CV cs.GR 73%

FSFSplatter: Build Surface and Novel Views with Sparse-Views within 2min

FSFSplatter: 用稀疏视图在2分钟内构建表面和新视角

Yibin Zhao, Yihan Pan, Jun Nan, Liwei Chen, Jianjun Yi

机构 * East China University of Science and Technology, Shanghai(东华大学上海理工学院) Shanghai Xiaoyuan Innovation Center, Shanghai(上海小元创新中心)

专题命中 三维重建 :Gaussian Splatting(abstract);novel view synthesis(abstract);分类 cs.CV、cs.GR

AI总结 本文提出FSFSplatter方法,通过端到端的密集高斯初始化、相机参数估计和几何增强场景优化,实现从自由稀疏图像快速重建表面和新视角,优于现有最先进方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.18795 2026-03-20 cs.CV cs.AI 57%

Perceptio: Perception Enhanced Vision Language Models via Spatial Token Generation

Perceptio:通过空间标记生成增强视觉语言模型

Yuchen Li, Amanmeet Garg, Shalini Chaudhuri, Rui Zhao, Garin Kessler

机构 * Amazon(亚马逊)

专题命中 三维重建 :spatial understanding(abstract);分类 cs.CV

AI总结 Perceptio通过生成2D和3D空间标记提升视觉语言模型的空间推理能力,结合语义分割和深度标记实现更精确的空间理解,提升多个基准测试性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20636 2026-03-20 cs.LG 50%

Image2Gcode: Image-to-G-code Generation for Additive Manufacturing Using Diffusion-Transformer Model

Image2Gcode: 基于扩散-变换器模型的图像到G-code生成用于增材制造

Ziyue Wang, Yayati Jadhav, Peter Pak, Amir Barati Farimani

机构 * Department of Materials Science and Engineering, Carnegie Mellon University, Pittsburgh, PA, USA(材料科学与工程系,卡内基梅隆大学,匹兹堡,PA,USA) Department of Mechanical Engineering, Carnegie Mellon University, Pittsburgh, PA, USA(机械工程系,卡内基梅隆大学,匹兹堡,PA,USA)

专题命中 三维重建 :3D reconstruction(abstract)

AI总结 本文提出Image2Gcode框架,通过图像直接生成可打印的G-code,省去CAD建模步骤,提升增材制造的易用性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏