arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

共收录 7719 信号源:cs.CV, cs.GR, cs.RO

1. 点云 7719 篇

2512.12410 2025-12-16 cs.CV cs.AI 57%

A Graph Attention Network-Based Framework for Reconstructing Missing LiDAR Beams

基于图注意力网络的缺失激光雷达束重建框架

Khalfalla Awedat, Mohamed Abidalrekab, Mohammad El-Yabroudi

机构 * Computer Information Technology Department, SUNY Morrisville(SUNY Morrisville 计算机信息科技系) Electrical and Computer Engineering Department, Portland State University(波特兰州立大学电气与计算机工程系) Electrical and Computer Engineering Department, Lawrence Technological University(劳伦斯技术大学电气与计算机工程系)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出基于图注意力网络的框架,通过点云几何数据恢复旋转激光雷达中因硬件老化等导致的垂直束丢失,实现高精度重建与快速推理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12377 2025-12-16 cs.RO 57%

INDOOR-LiDAR: Bridging Simulation and Reality for Robot-Centric 360 degree Indoor LiDAR Perception -- A Robot-Centric Hybrid Dataset

INDOOR-LiDAR: 联接仿真与现实以提升机器人中心的360度室内LiDAR感知 -- 一个机器人中心的混合数据集

Haichuan Li, Changda Tian, Panos Trahanias, Tomi Westerlund

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 INDOOR-LIDAR通过融合仿真与真实数据,为机器人感知研究提供了一个可扩展、真实且可重复的混合数据集,用于提升复杂室内环境中的感知能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13115 2025-12-16 cs.CV 57%

A Lightweight 3D Anomaly Detection Method with Rotationally Invariant Features

一种具有旋转不变特征的轻量3D异常检测方法

Hanzhe Liang, Jie Zhou, Can Gao, Bingyang Guo, Jinbao Wang, Linlin Shen

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) Shenzhen Audencia Financial Technology Institute, Shenzhen University(深圳大学深圳Audencia金融科技研究所) Audencia Nantes École de Management(Audencia南特商学院) School of Mathematics and Statistics, Changsha University of Science and Technology(长沙理工大学数学与统计学学院) Guangdong Provincial Key Laboratory of Intelligent Information Processing(广东省智能信息处理重点实验室) Software College, Northeastern University(东北大学软件学院) School of Artificial Intelligence, Shenzhen University(深圳大学人工智能学院)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出了一种轻量级3D异常检测方法,通过旋转不变特征框架提升检测性能,实验显示在多个数据集上均取得显著效果。

Comments Preprint. Accept by Pattern Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20991 2025-12-16 cs.CV 57%

MR-COSMO: Visual-Text Memory Recall and Direct CrOSs-MOdal Alignment Method for Query-Driven 3D Segmentation

MR-COSMO:一种用于查询驱动3D分割的视觉-文本记忆召回与直接跨模态对齐方法

Chade Li, Pengju Zhang, Yihong Wu

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 MR-COSMO通过视觉-文本记忆召回与直接跨模态对齐方法,在查询驱动的3D分割中实现几何与语义特征的精确融合,提升点云分割性能。

Comments Accepted by AAAI 2026. Copyright (c) 2026, Association for the Advancement of Artificial Intelligence (www.aaai.org). All rights reserved

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12013 2025-12-16 cs.CV cs.LG eess.IV 57%

Exploring Spatial-Temporal Representation via Star Graph for mmWave Radar-based Human Activity Recognition

通过星图探索毫米波雷达的人体活动识别中的时空表示

Senhao Gao, Junqing Zhang, Luoyu Mei, Shuai Wang, Xuyu Wang

机构 * School of Computer Science and Informatics, University of Liverpool(计算机科学与信息学系,利物浦大学) School of Computer Science and Engineering, Southeast University(计算机科学与工程系,东南大学) City University of Hong Kong(香港城市大学) Knight Foundation School of Computing and Information Sciences, Florida International University(Knight基金会计算与信息科学系,佛罗里达国际大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出基于星图和DDGNN的毫米波雷达人体活动识别方法,通过探索时空表示提升识别准确率,达到94.27%的分类精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11926 2025-12-16 cs.CV 57%

TransBridge: Boost 3D Object Detection by Scene-Level Completion with Transformer Decoder

TransBridge: 通过变换解码器进行场景级补全提升3D目标检测

Qinghao Meng, Chenming Wu, Liangjun Zhang, Jianbing Shen

机构 * School of Computer Science, Beijing Institute of Technology(北京理工大学计算机科学学院) Robotics and Autonomous Driving Lab (RAL), Baidu Research(百度研究机器人与自动驾驶实验室) State Key Laboratory of Internet of Things for Smart City, Department of Computer and Information Science, University of Macau(澳门大学智能城市物联网国家重点实验室,计算机与信息科学系)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 TransBridge通过变换解码器进行场景级补全,提升3D目标检测性能,实现稀疏区域特征增强和密集点云生成

Comments 12 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11894 2025-12-16 cs.CV cs.LG 57%

mmWEAVER: Environment-Specific mmWave Signal Synthesis from a Photo and Activity Description

mmWEAVER: 从照片和活动描述合成环境特定的毫米波信号

Mahathir Monjur, Shahriar Nirjon

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 mmWeaver通过隐式神经表示和超网络,高效生成环境特定的毫米波信号,提升活动识别和姿态估计性能。

Comments Accepted at the IEEE/CVF Winter Conference on Applications of Computer Vision 2026 (WACV 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11465 2025-12-15 cs.CV cs.LG 57%

DOS: Distilling Observable Softmaps of Zipfian Prototypes for Self-Supervised Point Representation

DOS: 通过Zipfian原型的可观察软图蒸馏实现自监督点表示

Mohamed Abdelsamad, Michael Ulrich, Bin Yang, Miao Zhang, Yakov Miron, Abhinav Valada

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 DOS通过Zipfian原型和可观察点软图蒸馏,提升3D点云自监督学习的语义分割和物体检测性能。

Comments AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09364 2025-12-11 cs.CV 57%

ASSIST-3D: Adapted Scene Synthesis for Class-Agnostic 3D Instance Segmentation

ASSIST-3D: 为无类别3D实例分割适应的场景合成

Shengchao Zhou, Jiehong Lin, Jiahui Liu, Shizhen Zhao, Chirui Chang, Xiaojuan Qi

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 ASSIST-3D通过异构对象选择、场景布局生成和点云构建,提升无类别3D实例分割的泛化能力。

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16372 2025-12-11 cs.RO 57%

Flow-Aided Flight Through Dynamic Clutters From Point To Motion

通过动态障碍物的飞行:从点到运动的流辅助飞行

Bowen Xu, Zexuan Yan, Minghao Lu, Xiyu Fan, Yi Luo, Youshen Lin, Zhiqiang Chen, Yeke Chen, Qiyuan Qiao, Peng Lu

机构 * Adaptive Robotic Controls Lab (ArcLab), Department of Mechanical Engineering, The University of Hong Kong(机械工程系,香港大学)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出了一种基于单LiDAR传感和强化学习的自主飞行系统,通过动态环境感知和动作生成方法,实现从点到运动的高效飞行控制。

Comments Accepted to IEEE Robotics and Automation Letters (RA-L), November, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11482 2025-12-11 cs.CV 57%

OpenConstruction: A Systematic Synthesis of Open Visual Datasets for Data-Centric Artificial Intelligence in Construction Monitoring

OpenConstruction: 一种系统性地合成开放视觉数据集的系统,用于数据驱动的建筑监控人工智能

Ruoxin Xiong, Yanyu Wang, Jiannan Cai, Kaijian Liu, Yuansheng Zhu, Pingbo Tang, Nora El-Gohary

机构 * College of Architecture & Environmental Design, Kent State University(建筑与环境设计学院,肯特州立大学) Bert S. Turner Department of Construction Management, Louisiana State University(建设管理伯特·S·Turner部门,路易斯安那州立大学) School of Civil & Environmental Engineering, and Construction Management, The University of Texas at San Antonio(土木与环境工程学院及建设管理学院,德克萨斯大学圣安东尼奥分校) Department of Civil, Environmental, and Ocean Engineering, Stevens Institute of Technology(土木、环境和海洋工程系,史蒂文斯理工学院) Department of Computing and Information Sciences, Rochester Institute of Technology(计算与信息科学系,罗切斯特理工学院) Department of Civil and Environmental Engineering, Carnegie Mellon University(土木与环境工程系,卡内基梅隆大学) Department of Civil and Environmental Engineering, University of Illinois at Urbana-Champaign(土木与环境工程系,伊利诺伊大学厄巴纳-香槟分校)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本研究系统性地整合了51个开放视觉数据集,构建了OpenConstruction开源目录,旨在提升建筑领域数据驱动的人工智能应用效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08653 2025-12-10 cs.RO 57%

A Sensor-Aware Phenomenological Framework for Lidar Degradation Simulation and SLAM Robustness Evaluation

一种面向传感器的现象学框架用于激光雷达退化仿真和SLAM鲁棒性评估

Doumegna Mawuto Koudjo Felix, Xianjia Yu, Zhuo Zou, Tomi Westerlund

机构 * Turku Intelligent Embedded and Robotic Systems (TIERS) Lab , University of Turku(图尔库智能嵌入与机器人系统实验室,图尔库大学) School of Information Science and Technology, Fudan University(信息科学与技术学院,复旦大学)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出一种面向传感器的现象学框架,用于模拟激光雷达退化并评估SLAM系统的鲁棒性,通过保留点云结构并应用多种退化方法,实现可控的仿真测试。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08330 2025-12-10 cs.CV 57%

PointDico: Contrastive 3D Representation Learning Guided by Diffusion Models

PointDico: 通过扩散模型引导的对比3D表示学习

Pengbo Li, Yiding Sun, Haozhe Cheng

机构 * International School Beijing University of Posts(国际学校 北京邮电大学) School of Software Engineering Xi'an Jiaotong University Xi'an, China(软件工程学院 西安交通大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 PointDico通过融合扩散模型和对比学习的方法,实现了3D表示学习的突破,达到了ScanObjectNN和ShapeNetPart上的新高精度。

Comments Accepted by IJCNN 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08135 2025-12-10 cs.CV 57%

CVP: Central-Peripheral Vision-Inspired Multimodal Model for Spatial Reasoning

CVP:基于中央-外围视觉的多模态模型用于空间推理

Zeyuan Chen, Xiang Zhang, Haiyang Xu, Jianwen Xie, Zhuowen Tu

机构 * UC San Diego(圣迭戈大学) Lambda, Inc(Lambda公司)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 CVP通过结合中央视觉和外围视觉的启发,提出了一种多模态模型,以提升复杂3D环境的空间推理能力。

Comments Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07599 2025-12-09 cs.CV 57%

Online Segment Any 3D Thing as Instance Tracking

在线分割三维物体作为实例跟踪

Hanshi Wang, Zijian Cai, Jin Gao, Yiwei Zhang, Weiming Hu, Ke Wang, Zhipeng Zhang

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院AutoLab) Anyverse Intelligence Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京超智能多模态信息安全重点实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 将在线3D分割重新定义为实例跟踪问题,通过时间信息传播和空间一致性学习提升具身智能体对环境的理解能力。

Comments NeurIPS 2025, Code is at https://github.com/AutoLab-SAI-SJTU/AutoSeg3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09620 2025-12-09 cs.CV cs.AI cs.CL 57%

Exploring the Potential of Encoder-free Architectures in 3D LMMs

探索无编码器架构在3D大语言模型中的潜力

Yiwen Tang, Zoey Guo, Zhuhao Wang, Ray Zhang, Qizhi Chen, Junli Liu, Delin Qu, Zhigang Wang, Dong Wang, Bin Zhao, Xuelong Li

机构 * Northwestern Polytechnical University(西北工业大学) Shanghai AI Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学) Tele AI

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出首个无编码器3D大语言模型ENEL,通过嵌入语义编码和分层几何聚合策略,在3D理解任务中达到与SOTA模型相当的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06517 2025-12-09 cs.RO 57%

Vision-Guided Grasp Planning for Prosthetic Hands in Unstructured Environments

面向无结构环境的假手视觉引导抓取规划

Shifa Sulaiman, Akash Bachhar, Ming Shen, Simon Bøgh

机构 * Control and Automation Section, Department of Electronics Systems, Aalborg University, Denmark(奥尔堡大学电子系统系自动化与自动化部门) Department of Mechanical Engineering, National Institute of Technology,Durgapur, India(印度德里加尔帕国家理工学院机械工程系) The Technical Faculty of IT(信息技术技术学院) Millimeter-Wave Systems, Department of Electronics Systems, Aalborg University, Denmark(毫米波系统,电子系统系,奥尔堡大学,丹麦)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出一种基于视觉引导的假手抓取算法,通过整合感知、规划和控制,实现灵活的抓取操作,并在仿真和实验中验证其在无结构环境中的适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09274 2025-12-09 cs.CV 57%

FLARES: Fast and Accurate LiDAR Multi-Range Semantic Segmentation

FLARES: 快速且准确的LiDAR多范围语义分割

Bin Yang, Alexandru Paul Condurache

机构 * Automated Driving Research, Robert Bosch GmbH(罗伯特博世集团自动化驾驶研究部) Institute for Signal Processing, University of Lübeck(吕贝克大学信号处理研究所)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 FLARES通过多范围图像训练和定制的数据增强技术,提升LiDAR语义分割的准确性和效率,实现mIoU提升和推理速度提升。

Comments The paper was accepted by WACV2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05482 2025-12-08 cs.CV 57%

Concept-based Explainable Data Mining with VLM for 3D Detection

基于概念的可解释数据挖掘与VLM用于3D检测

Mai Tsujimoto

机构 * The University of Tokyo(东京大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出基于概念的可解释数据挖掘方法,利用VLMs识别稀有物体以提升3D检测性能,减少标注负担并提高模型效果。

Comments 28 pages including appendix. Code: https://github.com/mm1129/concept_based_rare_detector_2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04563 2025-12-08 cs.CV 57%

COOPER: A Unified Model for Cooperative Perception and Reasoning in Spatial Intelligence

COOPER:一种用于空间智能中协作感知与推理的统一模型

Zefeng Zhang, Xiangzhao Hao, Hengzhu Tang, Zhenyu Zhang, Jiawei Sheng, Xiaodong Li, Zhenyang Li, Li Gao, Daiting Shi, Dawei Yin, Tingwen Liu

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Baidu Inc.(百度公司)

专题命中 点云 :spatial understanding(abstract);分类 cs.CV

AI总结 COOPER是一种统一的多模态大语言模型,通过整合深度和分割等辅助模态,提升空间感知与推理能力,实现空间智能的增强。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22143 2025-12-08 cs.CV 57%

3D Question Answering via only 2D Vision-Language Models

仅通过2D视觉-语言模型实现3D问答

Fengyun Wang, Sicheng Yu, Jiawei Wu, Jinhui Tang, Hanwang Zhang, Qianru Sun

机构 * Nanyang Technological University, Singapore Singapore Management University, Singapore National University of Singapore, Singapore Nanjing University of Science \& Technology, Nanjing, China

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出cdViews方法,通过仅使用2D视觉-语言模型,实现3D问答任务的高性能表现。

Comments ICML2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04996 2025-12-05 cs.CV 57%

A dynamic memory assignment strategy for dilation-based ICP algorithm on embedded GPUs

一种基于扩张的ICP算法在嵌入式GPU上的动态内存分配策略

Qiong Chang, Weimin Wang, Junpei Zhong, Jun Miyazaki

机构 * School of Coumputing Institute of Science(计算学院科学研究所) Dalian University of Technology(大连理工大学) University of Wollongong (College HK)(卧龙岗大学(香港学院)) School of Computing Institute of Science(计算学院科学研究所)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出了一种面向GPU的动态内存分配策略,用于优化VANICP算法的内存使用,实现97%的内存消耗减少,同时保持原有性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16848 2025-12-05 cs.CV 57%

A re-calibration method for object detection with multi-modal alignment bias in autonomous driving

面向自动驾驶的多模态对齐偏差校准方法

Zhihang Song, Dingyi Yao, Ruibo Ming, Lihui Peng, Danya Yao, Yi Zhang

机构 * Department of Automation Tsinghua University Beijing(自动化系清华大学北京)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出一种重新校准模型,通过语义分割和定制损失函数提升自动驾驶中多模态检测的鲁棒性和性能,应对校准偏差带来的影响。

Comments Accepted for publication in IST 2025. Official IEEE Xplore entry will be available once published

Journal ref 2025 IEEE International Conference on Imaging Systems and Techniques (IST)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03566 2025-12-04 cs.CV cs.MM 57%

GAOT: Generating Articulated Objects Through Text-Guided Diffusion Models

通过文本引导的扩散模型生成拟人化物体:GAOT

Hao Sun, Lei Fan, Donglin Di, Shaohui Liu

机构 * Harbin Institute of Technology(哈尔滨工业大学) University of New South Wales(新南威尔士大学)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 GAOT通过文本引导的扩散模型和超图学习,实现从文本提示到拟人化物体的生成,取得优于现有方法的性能。

Comments Accepted by ACM MM Asia2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.02394 2025-12-03 cs.CV 57%

Reproducing and Extending RaDelft 4D Radar with Camera-Assisted Labels

重现并扩展RaDelft 4D雷达与摄像头辅助标签

Kejia Hu, Mohammed Alsakabi, John M. Dolan, Ozan K. Tonguz

机构 * Department of Electrical and Computer Engineering, College of Engineering(电气与计算机工程系) The Robotics Institute, School of Computer Science(机器人研究所)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本研究重现并扩展RaDelft 4D雷达数据集,通过摄像头引导的标签生成方法实现雷达点云的高精度标注,并分析雾度对雷达标签性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05858 2025-12-03 cs.RO 57%

ViTaMIn-B: A Reliable and Efficient Visuo-Tactile Bimanual Manipulation Interface

ViTaMIn-B: 一种可靠且高效的视觉-触觉双臂操作接口

Chuanyu Li, Chaoyi Liu, Daotan Wang, Shuyu Zhang, Lusong Li, Zecui Zeng, Fangchen Liu, Jing Xu, Rui Chen

机构 * Tsinghua University(清华大学) University of California, Berkeley(加州大学伯克利分校) JD Explore Academy(JD探索学院) The Hong Kong Polytechnic University(香港理工大学)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 ViTaMIn-B通过DuoTact传感器和6自由度双臂位姿采集技术,实现了高效可靠的双臂操作任务数据采集。

Comments Project page: https://chuanyune.github.io/ViTaMIn-B_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07561 2025-12-03 cs.CV 57%

Alligat0R: Pre-Training Through Co-Visibility Segmentation for Relative Camera Pose Regression

Alligat0R:通过共视分割进行预训练以实现相对相机姿态回归

Thibaut Loiseau, Guillaume Bourmaud, Vincent Lepetit

机构 * LIGM, Ecole des Ponts, Univ. Gustave Eiffel, CNRS, France(LIGM,巴黎理工大学,埃菲尔大学,CNRS,法国) Univ. Bordeaux, CNRS, Bordeaux INP, IMS, UMR 5218, France(波尔多大学,CNRS,波尔多INP,IMS,UMR 5218,法国)

专题命中 点云 :3D reconstruction(abstract);分类 cs.CV

AI总结 Alligat0R通过共视分割预训练方法在相对相机姿态回归中优于CroCo

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01352 2025-12-02 cs.CV 57%

OpenBox: Annotate Any Bounding Boxes in 3D

OpenBox: 任意3D边界框的标注

In-Jae Lee, Mungyeom Kim, Kwonyoung Ryu, Pierre Musacchio, Jaesik Park

机构 * Seoul National University(首尔国立大学) POSTECH

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 OpenBox通过2D视觉基础模型实现无需自我训练的高质量3D边界框标注,提升自动驾驶中物体检测的准确性和效率。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10209 2025-12-02 cs.CV 57%

LiNeXt: Revisiting LiDAR Completion with Efficient Non-Diffusion Architectures

LiNeXt:重新审视基于高效非扩散架构的LiDAR补全

Wenzhe He, Xiaojun Chen, Ruiqi Wang, Ruihui Li, Huilong Pi, Jiapeng Zhang, Zhuo Tang, Kenli Li

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 LiNeXt通过高效非扩散架构实现快速准确的LiDAR点云补全,提升实时性能并减少计算开销。

Comments 18 pages, 13 figures, Accepted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27680 2025-12-02 cs.CV cs.AI cs.LG 57%

PETAR: Localized Findings Generation with Mask-Aware Vision-Language Modeling for PET Automated Reporting

PETAR:基于掩码感知的视觉-语言建模的局部发现生成用于PET自动报告

Danyal Maqbool, Changhee Lee, Zachary Huemann, Samuel D. Church, Matthew E. Larson, Scott B. Perlman, Tomas A. Romero, Joshua D. Warner, Meghan Lubner, Xin Tie, Jameson Merkow, Junjie Hu, Steve Y. Cho, Tyler J. Bradshaw

机构 * University of Wisconsin–Madison Department of Computer Sciences(威斯康星大学麦迪逊分校计算机科学系) University of Wisconsin–Madison Department Radiology(威斯康星大学麦迪逊分校放射学系) Microsoft(微软公司)

专题命中 点云 :3D vision(abstract);分类 cs.CV

AI总结 PETAR通过引入PETARSeg-11K数据集和PETAR-4B模型,实现基于掩码感知的3D PET自动报告生成,提升医学影像分析的精度与实用性。

详情

展开后加载摘要…

URL PDF HTML 收藏