arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-03-03 至 2026-03-03 共收录 55 信号源:cs.CV, cs.GR, cs.RO

1. Gaussian Splatting 14 篇

2508.19754 2026-03-03 cs.CV 70%

FastAvatar: Towards Unified and Fast 3D Avatar Reconstruction with Large Gaussian Reconstruction Transformers

FastAvatar: 向统一且快速的3D人像重建迈进:基于大高斯重建变换器

Yue Wu, Xuanhong Chen, Yufan Wu, Wen Li, Yuxi Lu, Kairui Feng

机构 * Tongji University(同济大学) Shanghai Innovation Institute(上海创新研究院) Shanghai Jiao Tong University(上海交通大学) AKool

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);3DGS(abstract);分类 cs.CV

AI总结 FastAvatar通过大高斯重建变换器实现快速且统一的3D人像重建,利用多种输入数据提升重建质量和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01491 2026-03-03 cs.CV cs.GR 62%

Radiometrically Consistent Gaussian Surfels for Inverse Rendering

辐射一致的高斯 Surfels 用于反演渲染

Kyu Beom Han, Jaeyoon Kim, Woo Jae Kim, Jinhwan Seo, Sung-eui Yoon

机构 * School of Computing(计算机学院) Korea Advanced Institute of Science and Technology(韩国科学技术院)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV、cs.GR

AI总结 RadioGS通过引入辐射一致性约束,结合高斯Surfels和二维高斯射线追踪,实现高效的反演渲染,提升对间接光照的建模精度。

Comments 9 pages, 6 figures, ICLR 2026 Oral paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02129 2026-03-03 cs.CV cs.AI 57%

LiftAvatar: Kinematic-Space Completion for Expression-Controlled 3D Gaussian Avatar Animation

LiftAvatar:基于运动空间的表达控制3D高斯人偶动画

Hualiang Wei, Shunran Jia, Jialun Liu, Wenhui Li

机构 * College of Computer Science and Technology, Jilin University(吉林大学计算机科学与技术学院) Institute of Artificial Intelligence of China Telecom(中国电信人工智能研究院)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 LiftAvatar通过多粒度表达控制和多参考条件机制,提升3D人偶动画的质量和可控性。

Comments 19 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01844 2026-03-03 cs.CV cs.AI 57%

CloDS: Visual-Only Unsupervised Cloth Dynamics Learning in Unknown Conditions

CloDS: 未知条件下的视觉-only 无监督布料动力学学习

Yuliang Zhan, Jian Li, Wenbing Huang, Wenbing Huang, Yang Liu, Hao Sun

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院) School of Engineering Science, University of Chinese Academy of Sciences(中国科学院大学工程科学学院)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 CloDS通过无监督学习从多视角视觉数据中学习布料动力学,解决未知条件下的动态模拟问题。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00145 2026-03-03 cs.CV cs.AI 57%

M-Gaussian: An Magnetic Gaussian Framework for Efficient Multi-Stack MRI Reconstruction

M-Gaussian:一种用于高效多堆栈MRI重建的磁性高斯框架

Kangyuan Zheng, Xuan Cai, Jiangqi Wang, Guixing Fu, Zhuoshuo Li, Yazhou Chen, Xinting Ge, Liangqiong Qu, Mengting Liu

机构 * School of Biomedical Engineering, Shenzhen Campus of Sun Yat-sen University(中山大学生物医学工程学院)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 M-Gaussian通过磁性高斯框架实现了高效多堆栈MRI重建,结合物理一致渲染与多分辨率训练,在FeTA数据集上达到40.31 dB PSNR并提升14倍速度。

Comments 15 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 点云 16 篇

2603.00412 2026-03-03 cs.CV 83%

PointAlign: Feature-Level Alignment Regularization for 3D Vision-Language Models

PointAlign:用于3D视觉-语言模型的特征级对齐正则化

Yuanhao Su, Shaofeng Zhang, Xiaosong Jia, Qi Fan

机构 * University of Science and Technology of China(中国科学技术大学) Fuzhou University(福州大学) Fudan University(复旦大学) Nanjing University(南京大学)

专题命中 点云 :3D vision(title,abstract);point cloud(abstract);分类 cs.CV

AI总结 PointAlign通过特征级对齐正则化提升3D视觉-语言模型的几何信息保留与任务性能

Comments CVPR 2026 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10492 2026-03-03 cs.CV eess.IV 83%

MFP3D: Monocular Food Portion Estimation Leveraging 3D Point Clouds

MFP3D:利用3D点云的单目食物分量估计

Jinge Ma, Xiaoyan Zhang, Gautham Vinod, Siddeshwar Raghavan, Jiangpeng He, Fengqing Zhu

机构 * Elmore Family School of Electrical and Computer Engineering, Purdue University, West Lafayette, USA(电气与计算机工程学院,普渡大学,西拉法叶,美国) College of Artificial Intelligence, Anhui University, Hefei, China(人工智能学院,安徽大学,合肥,中国)

专题命中 点云 :point cloud(title,abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 MFP3D通过单目图像和3D点云技术实现准确的食物分量估计,提升饮食监控的精度。

Comments 9th International Workshop on Multimedia Assisted Dietary Management, in conjunction with the 27th International Conference on Pattern Recognition (ICPR2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00870 2026-03-03 cs.CV cs.AI 79%

PPC-MT: Parallel Point Cloud Completion with Mamba-Transformer Hybrid Architecture

PPC-MT:基于Mamba-Transformer混合架构的并行点云补全

Jie Li, Shengwei Tian, Long Yu, Xin Ning

机构 * Xinjiang University(新疆大学) Institute of Semiconductors, Chinese Academy of Sciences(半导体研究所,中国科学院)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 PPC-MT通过混合Mamba-Transformer架构,提出并行点云补全方法,在效率与重建精度间取得平衡。

Comments Submitted to IEEE TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14218 2026-03-03 cs.CV 79%

Flexible-weighted Chamfer Distance: Enhanced Objective Function for Point Cloud Completion

灵活加权卡姆费距离:点云补全的增强目标函数

Jie Li, Shengwei Tian, Long Yu, Xin Ning

机构 * Xinjiang University(新疆大学) Institute of Semiconductors, Chinese Academy of Sciences(半导体研究所,中国科学院)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 FCD通过不对称加权策略提升点云补全的全局结构完整性,显著降低关键指标如DCD和EMD,增强点云的均匀性和结构完整性。

Comments Accepted by IEEE TPAMI 2026. This is the author's version of the work. \c{opyright} 2026 IEEE. Personal use of this material is permitted. Code is available at this https URL [https://github.com/Carroll-Li/FCD]

Journal ref IEEE Transactions on Pattern Analysis and Machine Intelligence, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01839 2026-03-03 cs.CV cs.RO 62%

LEAR: Learning Edge-Aware Representations for Event-to-LiDAR Localization

LEAR: 为事件到LiDAR定位学习边缘感知表示

Kuangyi Chen, Jun Zhang, Yuxi Hu, Yi Zhou, Friedrich Fraundorfer

机构 * Institute of Visual Computing, Graz University of Technology(视觉计算研究所,技术大学格拉茨) School of Artificial Intelligence and Robotics, Hunan University(人工智能与机器人学院,湖南大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 LEAR通过联合估计边缘结构和密集事件深度流场,解决事件与LiDAR对齐难题,提升GPS受限环境下的定位精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01074 2026-03-03 cs.CV 57%

Adaptive Augmentation-Aware Latent Learning for Robust LiDAR Semantic Segmentation

自适应增强感知潜在学习用于鲁棒激光雷达语义分割

Wangkai Li, Zhaoyang Li, Yuwen Pan, Rui Sun, Yujia Chen, Tianzhu Zhang

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 A3Point通过自适应增强感知潜在学习框架,有效缓解恶劣天气下的语义偏移问题,提升激光雷达语义分割的鲁棒性。

Comments Accepted by International Conference on Learning Representations (ICLR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00615 2026-03-03 cs.RO 57%

TGM-VLA: Task-Guided Mixup for Sampling-Efficient and Robust Robotic Manipulation

TGM-VLA:基于任务的混合学习用于高效且鲁棒的机器人操作

Fanqi Pu, Lei Jiang, Wenming Yang

机构 * Shenzhen International Graduate School, Tsinghua University, Shenzhen, China(清华大学深圳国际研究生院) The National and Local Co-Build Humanoid Robotics Innovation Center(国家级与地方共建人形机器人创新中心)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 TGM-VLA通过优化关键帧采样策略和引入颜色反转投影模块,提升机器人操作任务的效率和鲁棒性。

Comments 8 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00486 2026-03-03 cs.CV 57%

Random Wins All: Rethinking Grouping Strategies for Vision Tokens

随机胜出:重新思考视觉token的分组策略

Qihang Fan, Yuang Ai, Huaibo Huang, Ran He

机构 * MAIS & NLPR, Institute of Automation, Chinese Academy of Sciences, Beijing, China(自动化研究所,中国科学院,北京) School of Artificial Intelligence, University of Chinese Academy of Sciences, Beijing, China(人工智能学院,中国科学院大学,北京)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出随机分组策略,通过简单方法提升视觉token处理效率,实验显示其在多种任务中表现优异。

Comments Accepted by CVPR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00338 2026-03-03 cs.RO 57%

Layered Safety: Enhancing Autonomous Collision Avoidance via Multistage CBF Safety Filters

分层安全:通过多阶段CBF安全过滤器增强自主避障

Erina Yamaguchi, Ryan M. Bena, Gilbert Bahati, Aaron D. Ames

机构 * Caltech(卡内基梅隆大学)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出一种多阶段CBF安全过滤器,通过预测和实时安全过滤提升机器人动态避障的鲁棒性和安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15020 2026-03-03 cs.RO 57%

ISS Policy : Scalable Diffusion Policy with Implicit Scene Supervision

ISS政策:具有隐式场景监督的可扩展扩散策略

Wenlong Xia, Jinhao Zhang, Ce Zhang, Yaojia Wang, Huizhe Li, Youmin Gong, Jie Mei

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 ISS策略通过隐式场景监督模块提升机器人操作的性能和鲁棒性,实现高效、泛化能力强的3D视觉-运动扩散策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23926 2026-03-03 cs.CV 57%

Point-MoE: Large-Scale Multi-Dataset Training with Mixture-of-Experts for 3D Semantic Segmentation

点-MoE:通过混合专家进行大规模多数据集训练用于3D语义分割

Xuweiyi Chen, Wentao Zhou, Aruni RoyChowdhury, Zezhou Cheng

机构 * University of Virginia(弗吉尼亚大学) MathWorks

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 Point-MoE通过混合专家方法在无需数据集标签的情况下,实现大规模多数据集联合训练,提升3D语义分割性能。

Comments Project page: https://point-moe.cs.virginia.edu/

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15663 2026-03-03 cs.CV 57%

MSSPlace: Multi-Sensor Place Recognition with Visual and Text Semantics

MSSPlace: 多传感器位置识别与视觉和文本语义

Alexander Melekhin, Dmitry Yudin, Ilia Petryashin, Vitaly Bezuglyj

机构 * Intelligent Transport Laboratory, Moscow Institute of Physics and Technology(智能交通实验室,莫斯科物理技术学院) Artificial Intelligence Research Institute (AIRI)(人工智能研究机构(AIRI))

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 MSSPlace通过整合多传感器数据和视觉文本语义,提升位置识别性能,达到最先进的效果。

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17297 2026-03-03 cs.CV cs.AI 57%

Towards Camera Open-set 3D Object Detection for Autonomous Driving Scenarios

面向自动驾驶场景的相机开放集3D目标检测

Zhuolin He, Xinrun Li, Jiacheng Tang, Shoumeng Qiu, Wenfu Wang, Xiangyang Xue, Jian Pu

机构 * ZILab

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 OS-Det3D通过两阶段框架提升自动驾驶中相机3D目标检测器对未知对象的发现与识别能力,同时提升已知对象的检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09194 2026-03-03 math.GN 50%

Finding the Homology of Manifolds using Ellipsoids

使用椭球体寻找流形的同调

Sara Kalisnik, Davorin Lesnik

专题命中 点云 :point cloud(abstract)

AI总结 本文提出利用椭球体而非球体来计算流形的同调,从而降低样本密度要求并改进持续同调的构造方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.00470 2026-03-03 math.PR 50%

Geometric and Probabilistic Limit Theorems in Topological Data Analysis

几何与概率极限定理在拓扑数据分析中的应用

Sara Kalisnik, Christian Lehn, Vlada Limic

专题命中 点云 :point cloud(abstract)

AI总结 本文提出了一种分析随机有限点云的概率框架,证明了条形码在点数趋于无穷时的收敛性,并探讨了在紧致流形上均匀采样的定量收敛问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14994 2026-03-03 math.NA cs.NA hep-ph math-ph math.MP math.OC physics.comp-ph 50%

Optimal alignment of Lorentz orientation and generalization to matrix Lie groups

洛伦兹取向的最优对齐及其对矩阵李群的推广

Congzhou M Sha

专题命中 点云 :point cloud(abstract)

AI总结 本文提出两种方法解决四向量在不同参考系间的最优洛伦兹变换问题,并推广到其他矩阵李群的向量对齐

Comments 10 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 新视角合成 1 篇

2509.04932 2026-03-03 cs.CV 79%

UniView: Enhancing Novel View Synthesis From A Single Image By Unifying Reference Features

UniView: 通过统一参考特征从单张图像增强新颖视角合成

Haowang Cui, Rui Chen, Jiaze Wang, Tao Guo, Zheng Qin

机构 * Tianjin University(天津大学)

专题命中 新视角合成 :novel view synthesis(title,abstract);分类 cs.CV

AI总结 UniView通过统一参考特征提升单张图像新颖视角合成性能,采用检索增强系统和多模态大语言模型选择参考图像,并结合解耦三重注意力机制提高合成效果。

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 3D生成 1 篇

2603.01594 2026-03-03 cs.CV 79%

Preference Score Distillation: Leveraging 2D Rewards to Align Text-to-3D Generation with Human Preference

偏好分数蒸馏:利用2D奖励对齐文本到3D生成

Jiaqi Leng, Shuyuan Tu, Haidong Cao, Sicheng Xie, Daoguo Dong, Zuxuan Wu, Yu-Gang Jiang

机构 * Fudan University(复旦大学)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 本文提出偏好分数蒸馏方法,利用2D奖励模型实现文本到3D生成的人类偏好对齐,无需3D训练数据,提升生成质量与扩展性。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 空间理解 1 篇

2510.23607 2026-03-03 cs.CV 70%

Concerto: Joint 2D-3D Self-Supervised Learning Emerges Spatial Representations

Concerto:联合2D-3D自监督学习产生空间表示

Yujia Zhang, Xiaoyang Wu, Yixing Lao, Chengyao Wang, Zhuotao Tian, Naiyan Wang, Hengshuang Zhao

机构 * The University of Hong Kong(香港大学) The Chinese University of Hong Kong(香港中文大学) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))

专题命中 空间理解 :point cloud(abstract);spatial understanding(abstract);分类 cs.CV

AI总结 Concerto通过联合2D-3D自监督学习,产生更连贯的信息空间表示,优于现有SOTA模型并在多个基准上取得新成就。

Comments NeurIPS 2025, produced by Pointcept, project page: https://pointcept.github.io/Concerto

Journal ref Neural Information Processing Systems 2025

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 其他3D视觉 1 篇

2603.00545 2026-03-03 cs.CV 79%

Multiple Inputs and Mixwd data for Alzheimer's Disease Classification Based on 3D Vision Transformer

基于3D视觉变换器的阿尔茨海默病分类的多输入和混合数据

Juan A. Castro-Silva, Maria N. Moreno Garcia, Diego H. Peluffo-Ordoñez

专题命中 其他3D视觉 :3D vision(title,abstract);分类 cs.CV

AI总结 本研究提出MIMD-3DVT方法,通过多输入和混合数据提升阿尔茨海默病分类准确率至97.14%

详情

展开后加载摘要…

URL PDF HTML 收藏