arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2026-07-03 至 2026-07-03 共收录 37 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 8 篇

2607.01851 2026-07-03 cs.CV 新提交 79%

Geometric Foundation Model Distillation for Efficient Lunar 3D Reconstruction

几何基础模型蒸馏用于高效月球三维重建

Clémentine Grethen, Florient Chouteau, Géraldine Morin, Simone Gasparini

机构 * IRIT, University of Toulouse(图卢兹大学IRIT研究所) Airbus Defence and Space(空中客车防务与航天公司)

专题命中 三维重建 :3D reconstruction(title,abstract);分类 cs.CV

AI总结 针对硬件受限场景,通过知识蒸馏将MASt3R模型压缩7倍,保留大部分重建精度,提出SVD初始化方法提升训练稳定性,并揭示编码器类型、蒸馏策略等关键因素。

Comments Accepted to ECCV 2026, code can be accessed via https://clementinegrethen.github.io/publications/ECCV.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01753 2026-07-03 cs.CV q-bio.QM 新提交 70%

The Turning Point of 3D Plant Phenotyping: 3D Foundation Models Enable Minute-to-Second Cross-Crop Reconstruction and Beyond

3D植物表型分析的转折点:3D基础模型实现分钟到秒级的跨作物重建及更多

Hanyue Jia, Wei Zhou, Wenbo Zhou, Yanan Li, Hao Lu, Tingting Wu

机构 * Northwest A&F University(西北农林科技大学) Huazhong University of Science and Technology(华中科技大学) Wuhan Institute of Technology(武汉工程大学)

专题命中 三维重建 :Gaussian Splatting(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 提出基于3D基础模型的跨作物表型分析框架,用前馈几何恢复替代COLMAP,结合3D高斯泼溅实现少视图重建,将重建时间从6.52分钟降至1.58秒,保持高质量与高精度。

Comments 39 pages, 6 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02515 2026-07-03 cs.CV 新提交 57%

PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimation

PointDiT: 用于单目几何估计的像素空间扩散

Haofei Xu, Rundi Wu, Philipp Henzler, Nikolai Kalischek, Michael Oechsle, Fabian Manhardt, Marc Pollefeys, Andreas Geiger, Federico Tombari, Michael Niemeyer

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 提出基于纯ViT的像素空间扩散Transformer,直接在原始3D点图块上操作,无需复杂架构和潜在空间压缩,在单目几何估计中超越潜在扩散模型。

Comments ICML 2026. Project page: https://haofeixu.github.io/pointdit/

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01454 2026-07-03 cs.RO 新提交 57%

SE(2) Navigation Mesh

SE(2)导航网格

Shuyang Shi, Kaixian Qu, Changan Chen, Ines Kast, Yuntao Ma, Marco Hutter

机构 * Robotic Systems Lab, ETH Zurich(苏黎世联邦理工学院机器人系统实验室)

专题命中 三维重建 :point cloud(abstract);分类 cs.RO

AI总结 提出SE(2)导航网格,通过编码偏航依赖的可通行性,支持非圆形机器人在复杂多层环境中的高效路径规划,并开发了A*-拉绳-A*分层路径搜索策略。

Comments Project page: https://se2-navmesh.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01367 2026-07-03 cs.MA cs.RO 新提交 57%

Simulation Based Reward Function Validation for Multi-Agent On Orbit Inspection

基于仿真的多智能体在轨检测奖励函数验证

Patrick Quinn, Bala Prenith Reddy Gopu, George M. Nehma, Madhur Tiwari

机构 * Department of Aerospace, Physics and Space Sciences(航空航天、物理与空间科学系)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.RO

AI总结 提出一种基于多智能体强化学习的广义奖励函数,通过分析在轨物体三维重建来优化检测任务,使智能体自主控制图像采集时机。

Comments 13 pages, 6 figures. This submission integrates a published correction made to the original manuscript. The DOIs for both the original manuscript as well as the correction are provided

Journal ref AIAA SCITECH 2026 Forum, AIAA Paper 2026-2042, 2026; Correction: AIAA SCITECH 2026 Forum, AIAA Paper 2026-2042.c1, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.11326 2026-07-03 cs.CV 新提交 57%

DarkVGGT: Seeing Through Darkness Using Thermal Geometry without Daylight Tax

DarkVGGT: 利用热几何在黑暗中透视,无需日光代价

Minseong Kweon, Wenyuan Zhao, Nuo Chen, Lulin Liu, Huiwen Han, Zihao Zhu, Srinivas Shakkottai, Chao Tian, Zhiwen Fan

机构 * University of Minnesota(明尼苏达大学) Texas A&M University(德克萨斯农工大学) Stanford University(斯坦福大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 提出DarkVGGT,一种RGB-T前馈几何框架,通过物理感知热建模实现低光照场景下的鲁棒3D估计,引入热分解和几何共享路由模块,在退化RGB条件下保持精度。

Comments Project Page: https://darkvggt.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28896 2026-07-03 cs.CV 版本更新 57%

Fisheye3R: Adapting Unified 3D Feed-Forward Foundation Models to Fisheye Lenses

Fisheye3R:将统一的3D前馈基础模型适应于鱼眼镜头

Ruxiao Duan, Erin Hong, Dongxu Zhao, Eric Turner, Alex Wong, Yunwen Zhou

机构 * Yale University(耶鲁大学) Google XR(谷歌XR)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 本文提出Fisheye3R框架,通过灵活学习方案在不退化性能的前提下,使多视角3D重建模型适应高径向畸变的鱼眼图像。

Comments European Conference on Computer Vision 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01204 2026-07-03 cs.CV 版本更新 57%

TabletopGen: Tabletop Scene Generation and Interactive Simulation for Robotic Manipulation

TabletopGen: 桌面场景生成与机器人操作的交互式仿真

Ziqian Wang, Yonghao He, Licheng Yang, Wei Zou, Hongxuan Ma, Liu Liu, Wei Sui, Yuxin Guo, Hu Su

机构 * School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences(中国科学院多模态人工智能系统国家重点实验室) D-Robotics Horizon Robotics

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 提出TabletopGen,一种无需训练的全自动桌面场景生成与交互仿真引擎,通过实例提取、位姿对齐和物理仿真生成可交互场景,用于机器人操作数据合成。

Comments Project page: https://d-robotics-ai-lab.github.io/TabletopGen.project/

详情

展开后加载摘要…

URL PDF HTML 收藏

2. Gaussian Splatting 10 篇

2607.01556 2026-07-03 cs.CV 新提交 92%

Mind the Gap: Standard 3DGS Evaluation Primarily Measures Near-Trajectory Interpolation

注意差距:标准 3DGS 评估主要衡量近轨迹插值

Gaoxiang Jia, Vikram Appia

机构 * Advanced Micro Devices, Inc.(超威半导体公司)

专题命中 Gaussian Splatting :3DGS(title,title_cn);NeRF(abstract,abstract_cn);Gaussian Splatting(abstract);分类 cs.CV

AI总结 标准 3DGS 评估采用隔帧留出法,实际衡量近轨迹插值而非空间泛化。本文提出匹配计数协议,发现插值与外推之间存在 3~12dB 的稳定差距,该差距跨表示族存在,主要由几何代理分量主导,并建议采用空间留出基准。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02099 2026-07-03 cs.CV 新提交 89%

X-Splat: Gaussian Splatting for 3D CBCT Generation from Single Panoramic Radiograph

X-Splat:基于高斯泼溅的从单张全景X光片生成3D CBCT

Tomasz Szczepański, Szymon Płotka, Michal K. Grzeszczyk, Tomasz Trzciński, Arkadiusz Sitek

机构 * Sano Centre for Computational Medicine(桑诺计算医学中心) Jagiellonian University(雅盖隆大学) Warsaw University of Technology(华沙技术大学) Research Institute IDEAS(IDEAS研究 institute) Harvard Medical School(哈佛医学院)

专题命中 Gaussian Splatting :NeRF(summary_cn,abstract);Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 提出X-Splat,首个利用高斯泼溅从单张全景X光片生成CBCT样3D牙科体积的方法,通过已知采集几何初始化可学习高斯原语,结合Beer-Lambert重投影和多视角训练监督,优于NeRF和GAN基线。

Comments 19 pages, 6 figures, including appendix. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01708 2026-07-03 cs.CV 新提交 85%

Consistent Scene Understanding in 3D Gaussian Splatting via Multi-Cue Mask Refinement

基于多线索掩码精炼的三维高斯泼溅中一致场景理解

Hyunjoon Park, Donghyeon Cho

机构 * Department of Computer Science, Hanyang University(汉阳大学计算机科学系)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract,abstract_cn);分类 cs.CV

AI总结 提出一种多线索掩码精炼框架,通过融合语义、几何和结构先验,实现跨视图一致的2D实例掩码,从而优化3D高斯泼溅特征场,提升3D实例分割和下游编辑任务的稳定性与一致性。

Comments Accepted at ICPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23998 2026-07-03 cs.CV 83%

Improved 3D Gaussian Splatting of Unknown Spacecraft Structure Using Space Environment Illumination Knowledge

利用空间环境照明知识改进未知航天器结构的3D高斯点扩散

Tae Ha Park, Simone D'Amico

机构 * Nara Space Technology Inc.(那拉航天科技公司) Dept. of Aeronautics & Astronautics, Stanford University(航空航天系,斯坦福大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 本文提出一种新流程,通过捕捉 rendezvous 和 proximity operations (RPO) 期间的图像序列,恢复未知目标航天器的3D结构。利用太阳位置先验知识提升3DGS的光照质量,以提高下游姿态估计任务的性能。

Comments Presented at 2025 IEEE International Conference on Space Robotics (iSpaRo)

Journal ref 2025 International Conference on Space Robotics (iSpaRo)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00595 2026-07-03 cs.CV 新提交 79%

GADA: Geometry-Aware Deformable Aggregation for Image-Based Gaussian Splatting

GADA: 基于图像的高斯泼溅的几何感知可变形聚合

Siwoo Lim, Sunjae Yoon, Gwanhyeong Koo, Chang D. Yoo

机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院) Chung-Ang University(中央大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 提出几何感知可变形聚合(GADA),通过可变形偏移迭代校正空间错位,并引入隐式置信加权机制抑制不可靠证据,在保持高频细节的同时实现2.13倍FPS提升。

Comments ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19703 2026-07-03 cs.CV eess.IV 79%

High-Quality Spatial Reconstruction and Orthoimage Generation Using Efficient 2D Gaussian Splatting

利用高效2D高斯散射实现高质量空间重建和正射影像生成

Qian Wang, Zhihao Zhan, Jialei He, Zhituo Tu, Jie Yuan

机构 * TopXGun Robotics(TopXGun机器人)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 本文提出基于2D高斯散射的高效方法,实现高质量空间重建和正射影像生成,无需传统DSM和遮挡检测,提升复杂地形和细长结构的渲染质量与效率。

Journal ref Signal, Image and Video Processing, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01628 2026-07-03 cs.CV 新提交 77%

Online Segment 3D Gaussians via Launching Virtual Drones

通过发射虚拟无人机在线分割3D高斯体

Liwei Liao, Rongjie Wang, Ronggang Wang

机构 * Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology, Shenzhen Graduate School, Peking University(广东省超高清沉浸式媒体技术重点实验室,北京大学深圳研究生院) Pengcheng Laboratory(鹏城实验室) Peking University(北京大学)

专题命中 Gaussian Splatting :3DGS(abstract,abstract_cn);Gaussian Splatting(abstract);分类 cs.CV

AI总结 提出SAGO框架,通过虚拟无人机将3D分割转化为在线最佳视角规划任务,无需预设置即可在亚秒内从3D高斯体中提取干净3D资产。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01521 2026-07-03 cs.CV 版本更新 74%

MiraGe: Editable 2D Images using Gaussian Splatting

MiraGe: 使用高斯溅射的可编辑2D图像

Joanna Waczyńska, Tomasz Szczepanik, Piotr Borycki, Sławomir Tadeja, Thomas Bohné, Przemysław Spurek

专题命中 Gaussian Splatting :Gaussian Splatting(title);分类 cs.CV

AI总结 提出MiraGe方法,利用镜面反射在3D空间中感知2D图像,并通过平面控制的高斯函数实现精确编辑,提升渲染质量并支持物理仿真修改。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24009 2026-07-03 cs.GR cs.AI cs.CV cs.LG cs.RO 版本更新 67%

Learning 3D-Gaussian Simulators from RGB Videos

从RGB视频学习3D高斯模拟器

Mikel Zhobro, Andreas René Geist, Georg Martius

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV、cs.GR、cs.RO

AI总结 提出3DGSim,一种从多视角RGB视频直接学习物理交互的3D模拟器,统一了3D场景重建、粒子动力学预测和视频合成,能泛化到未见过的多体交互和新场景编辑。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29453 2026-07-03 cs.CV cs.AI cs.GR 版本更新 62%

Resonant Brane Splatting for Arbitrary-Scale Super-Resolution

用于任意尺度超分辨率的共振膜片溅射

Giulio Federico, Giuseppe Amato, Claudio Gennaro, Fabio Carrara, Marco Di Benedetto

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV、cs.GR

AI总结 提出共振膜片溅射(RBS)框架,用Brane基元替代高斯溅射,通过高斯-埃尔米特模式捕获高频细节,实现任意尺度超分辨率的高质量与高效渲染。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 点云 10 篇

2607.01272 2026-07-03 cs.GR cs.AI cs.CV cs.DC cs.LG 新提交 81%

Benchmarking Federated Learning and Knowledge Distillation for Point Cloud Classification

联邦学习与知识蒸馏在点云分类中的基准测试

Aizierjiang Aiersilan

机构 * University of Macau(澳门大学)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV、cs.GR

AI总结 针对隐私敏感和资源受限场景,联合评估联邦学习与知识蒸馏在3D点云分类中的性能,发现极端非独立同分布标签偏移下联邦学习性能下降,蒸馏可压缩模型且避免标签泄露问题。

Comments We are pleased to announce that this paper has been accepted by the 19th European Conference on Computer Vision (ECCV 2026). We appreciate the valuable feedback from the reviewers and look forward to sharing our findings with the community

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02225 2026-07-03 hep-ex 新提交 78%

Heavy-Flavor Electron Classification Using Hadronic Environment as Point Cloud

使用强子环境作为点云的重味电子分类

Jingyu Zhang, Wanbing He, Long Ma

专题命中 点云 :point cloud(title,abstract)

AI总结 将强子环境表示为点云,利用集合机器学习架构区分粲和底起源电子,发现性能受限于强子结构内在相似性,在40%效率下纯度约80%,优于传统BDT基线。

Comments 8 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02074 2026-07-03 cs.CV 新提交 57%

Comprehensive Robustness Analysis of LiDAR-based 3D Object Detection in Autonomous Driving

自动驾驶中基于LiDAR的3D物体检测的全面鲁棒性分析

Adwait Chandorkar, Kai Krink, Yerdana Maulenbay, Hasan Tercan, Tobias Meisen

机构 * Institute for TMDT, University of Wuppertal, Germany(TMDT研究所,乌尔姆大学,德国) IKB Faculty of Science, University of British Columbia, Canada(科学学院,不列颠哥伦比亚大学,加拿大)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 针对LiDAR-only 3D物体检测模型对抗鲁棒性研究不足的问题,提出一个综合评估框架,通过结构因子(点云密度、定位)和预测因子(误分类、定位误差、自车距离)分析,发现高容量体素检测器比支柱检测器更易受结构化坐标扰动,非锚点检测器鲁棒性差。

Comments Accepted at ECCV 2026 main

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02018 2026-07-03 cs.CV 新提交 57%

UnderOneFacade: Worldwide Facade Semantic Segmentation Benchmark Dataset

UnderOneFacade:全球立面语义分割基准数据集

Yi Wang, Fan Wang, Prabin Gyawali, Ziyang Xu, Anna Klimkowska, Yixiong Jing, Wanru Yang, Filip Biljecki, Christoph Holst, Benjamin Busam, Brian Sheil, Olaf Wysocki

机构 * Technical University of Munich(慕尼黑工业大学) University of Cambridge(剑桥大学) University of Nottingham(诺丁汉大学) National University of Singapore(新加坡国立大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 针对现有立面解析数据集地理范围窄、语义不一致的问题,提出全球最大跨国家、跨洲的3D立面基准UnderOneFacade,包含27亿标注点,评估多种模型并揭示其跨域性能下降,为鲁棒可迁移的3D分割模型提供基准。

Comments accepted by ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01827 2026-07-03 cs.CV 新提交 57%

C2E: Boosting Ego-Only 3D Object Detection via Multi-Teacher Contrastive Knowledge Distillation

C2E: 通过多教师对比知识蒸馏提升仅自我3D目标检测

Jinlong Wang, Xun Huang, Qiming Xia, Shijia Zhao, Chenglu Wen

机构 * Fujian Key Laboratory of Urban Intelligent Sensing and Computing, Xiamen University(福建省城市智能感知与计算重点实验室,厦门大学) Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(多媒体可信感知与高效计算教育部重点实验室,厦门大学) Zhongguancun Academy(中关村学院)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 提出C2E范式,通过多对一智能体对比知识蒸馏框架M2S,结合多级特征增强、辅助点云重建和多教师对比蒸馏,在无通信成本下提升仅自我3D检测性能,在多个数据集上验证有效性。

Comments 18 pages, 8figures

Journal ref ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.31918 2026-07-03 cs.CV 新提交 57%

DriveWeaver: Point-Conditioned Video Inpainting for Controllable Vehicle Insertion in Autonomous Driving Simulation

DriveWeaver: 基于点云条件的视频修复用于自动驾驶仿真中的可控车辆插入

Junzhe Jiang, Zipei Ma, Zijie Pan, Li Zhang

机构 * School of Data Science, Fudan University(复旦大学数据科学学院) Shanghai Innovation Institute(上海创新研究院)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 提出DriveWeaver框架,通过点云条件视频修复实现自动驾驶仿真中可控车辆插入,解决光照不一致和3D资产依赖问题,支持大规模场景增强。

Comments Accepted at ECCV 2026, Project Page: https://github.com/LogosRoboticsGroup/DriveWeaver

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09604 2026-07-03 cs.CV 版本更新 57%

DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition

DAP:面向异构毫米波的点网络用于人类动作识别

Jiaying Lin, Shiman Wu, Jinfu Liu, Can Wang, Mengyuan Liu

机构 * Peking University(北京大学) Huazhong University of Science and Technology(华中科技大学) DJI Technology Company Ltd.(大疆技术创新有限公司) Christian-Albrechts-Universität zu Kiel(基尔大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出DAP-Net,通过异构雷达源的跨模态对齐和D2R模块实现点云增强,提升动作识别的鲁棒性与准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14421 2026-07-03 cs.RO 版本更新 57%

BIEVR-LIO: Robust LiDAR-Inertial Odometry through Bump-Image-Enhanced Voxel Maps

BIEVR-LIO:通过凸起图像增强体素地图实现鲁棒的激光雷达惯性里程计

Patrick Pfreundschuh, Turcan Tuna, Cedric Le Gentil, Roland Siegwart, Cesar Cadena, Helen Oleynikova

机构 * 1 Autonomous Systems Lab, ETH Z\"urich, Switzerland, 2 Robotic Systems Lab, ETH Z\"urich, Switzerland, 3 Mobile Robotics Lab, ETH Z\"urich, Switzerland -1.5mm

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出BIEVR-LIO方法,通过高分辨率体素地图表示和信息导向点采样策略,提升在信息稀疏环境下的激光雷达惯性里程计鲁棒性,并在多种场景中实现优异性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10816 2026-07-03 cs.CV 版本更新 57%

Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders

基于掩码自编码器的遮挡感知3D手-物体姿态估计

Hui Yang, Wei Sun, Jian Liu, Jin Zheng, Jian Xiao, Ajmal Mian

机构 * National Engineering Research Center for Robot Visual Perception and Control Technology, School of Artificial Intelligence and Robotics, Hunan University(国家机器人视觉感知与控制技术工程研究中心,人工智能与机器人学院,湖南大学) School of Architecture and Art, Central South University(建筑与艺术学院,中南大学) Department of Computer Science and Software Engineering, The University of Western Australia(计算机科学与软件工程系,西澳大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 提出HOMAE方法,通过目标聚焦掩码策略和隐式SDF与显式点云融合,解决手-物体交互中的严重遮挡问题,在DexYCB和HO3Dv2上达到最优性能。

Comments IEEE Transactions on Multimedia 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04357 2026-07-03 math.AT 版本更新 50%

Circular Max-Flow for Periodic Data via Reeb Graphs

基于Reeb图的周期数据环形最大流

Matteo Pegoraro, Lisbeth Fajstrup

专题命中 点云 :point cloud(abstract)

AI总结 针对周期性边界条件数据,利用Reeb图将空间几何简化为有向一维隧道网络,并引入容量约束定义环形最大流,通过线性优化求解,无需指定进出口。

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 新视角合成 4 篇

2604.01761 2026-07-03 cs.CV 版本更新 77%

Control-DINO: Feature Space Conditioning for Controllable Image-to-Video Diffusion

Control-DINO:用于可控图像到视频扩散的特征空间条件

Edoardo A. Dominici, Thomas Deixelberger, Konstantinos Vardis, Markus Steinberger

机构 * Huawei Technologies, Switzerland(华为技术(瑞士)) Huawei Technologies, Austria(华为技术(奥地利)) Graz University of Technology, Austria(格拉茨技术大学)

专题命中 新视角合成 :point cloud(abstract);novel view synthesis(abstract);3D generation(abstract);分类 cs.CV

AI总结 本文提出Control-DINO,通过解耦外观与其他特征,实现对视频扩散模型的可控生成,提升生成渲染的可控性。

Comments ECCV 2026 - Project Page https://dedoardo.github.io/projects/control-dino/

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.02372 2026-07-03 cs.CV 新提交 74%

Learning Spectral and Polarimetric Clues for One-to-Multimodal Novel View Synthesis

学习光谱和偏振线索用于一对多模态新视角合成

Federico Lincetto, Gianluca Agresti, Mattia Rossi, Piergiorgio Sartor, Pietro Zanuttigh

机构 * MEDIA Lab(媒体实验室) University of Padova(帕多瓦大学) Sony EUISPC(索尼欧洲影像处理中心) Sony Semiconductor Solutions Europe(索尼半导体解决方案欧洲分公司)

专题命中 新视角合成 :novel view synthesis(title);分类 cs.CV

AI总结 提出SPoILeR方法,通过多模态预训练学习模态间相关性,在仅RGB监督下实现未见场景的多模态新视角渲染。

Comments Accepted at ECCV 2026. Project page: https://medialab.dei.unipd.it/paper_data/SPoILeR/

详情

展开后加载摘要…

URL PDF HTML 收藏