arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2507.07620 2025-09-22 cs.CV

ViLU: Learning Vision-Language Uncertainties for Failure Prediction

Marc Lafon, Yannis Karmim, Julio Silva-Rodríguez, Paul Couairon, Clément Rambour, Raphaël Fournier-Sniehotta, Ismail Ben Ayed, Jose Dolz, Nicolas Thome

Journal ref International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01064 2025-09-22 cs.CV cs.AI cs.LG cs.MM eess.IV

FLOAT: Generative Motion Latent Flow Matching for Audio-driven Talking Portrait

Taekyung Ki, Dongchan Min, Gyeongsu Chae

机构 * KAIST(韩国科学技术院) DeepBrain AI Inc.(DeepBrain AI公司)

Comments ICCV 2025. Project page: https://deepbrainai-research.github.io/float/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15224 2025-09-19 cs.CV

Depth AnyEvent: A Cross-Modal Distillation Paradigm for Event-Based Monocular Depth Estimation

Luca Bartolomei, Enrico Mannocci, Fabio Tosi, Matteo Poggi, Stefano Mattoccia

机构 * Advanced Research Center on Electronic System (ARCES)(电子系统先进研究中心) Department of Computer Science and Engineering (DISI)(计算机科学与工程系) University of Bologna(博洛尼亚大学)

Comments ICCV 2025. Code: https://github.com/bartn8/depthanyevent/ Project Page: https://bartn8.github.io/depthanyevent/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11277 2025-09-19 cs.CV cs.LG

Probing the Representational Power of Sparse Autoencoders in Vision Models

Matthew Lyle Olson, Musashi Hinck, Neale Ratzlaff, Changbai Li, Phillip Howard, Vasudev Lal, Shao-Yen Tseng

机构 * Oracle Intel Labs(英特尔实验室) Oregon State University(俄勒冈州立大学) Thoughtworks

Comments ICCV 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22422 2025-09-19 cs.CV

Gradient Distance Function

Hieu Le, Federico Stella, Benoit Guillard, Pascal Fua

机构 * UNC-Charlotte(北卡罗来纳大学夏洛特分校) EPFL(苏黎世联邦理工学院)

Comments ICCV 2025 - Wild3D workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14142 2025-09-18 cs.CV

MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook

Peng Xu, Shengwu Xiong, Jiajun Zhang, Yaxiong Chen, Bowen Zhou, Chen Change Loy, David A. Clifton, Kyoung Mu Lee, Luc Van Gool, Ruiming He, Ruilin Yao, Xinwei Long, Jirui Huang, Kai Tian, Sa Yang, Yihua Shao, Jin Feng, Yue Zhong, Jiakai Zhou, Cheng Tang, Tianyu Zou, Yifang Zhang, Junming Liang, Guoyou Li, Zhaoxiang Wang, Qiang Zhou, Yichen Zhao, Shili Xiong, Hyeongjin Nam, Jaerin Lee, Jaeyoung Chung, JoonKyu Park, Junghun Oh, Kanggeon Lee, Wooseok Lee, Juneyoung Ro, Turghun Osman, Can Hu, Chaoyang Liao, Cheng Chen, Chengcheng Han, Chenhao Qiu, Chong Peng, Cong Xu, Dailin Li, Feiyu Wang, Feng Gao, Guibo Zhu, Guopeng Tang, Haibo Lu, Han Fang, Han Qi, Hanxiao Wu, Haobo Cheng, Hongbo Sun, Hongyao Chen, Huayong Hu, Hui Li, Jiaheng Ma, Jiang Yu, Jianing Wang, Jie Yang, Jing He, Jinglin Zhou, Jingxuan Li, Josef Kittler, Lihao Zheng, Linnan Zhao, Mengxi Jia, Muyang Yan, Nguyen Thanh Thien, Pu Luo, Qi Li, Shien Song, Shijie Dong, Shuai Shao, Shutao Li, Taofeng Xue, Tianyang Xu, Tianyi Gao, Tingting Li, Wei Zhang, Weiyang Su, Xiaodong Dong, Xiao-Jun Wu, Xiaopeng Zhou, Xin Chen, Xin Wei, Xinyi You, Xudong Kang, Xujie Zhou, Xusheng Liu, Yanan Wang, Yanbin Huang, Yang Liu, Yang Yang, Yanglin Deng, Yashu Kang, Ye Yuan, Yi Wen, Yicen Tian, Yilin Tao, Yin Tang, Yipeng Lin, Yiqing Wang, Yiting Xi, Yongkang Yu, Yumei Li, Yuxin Qin, Yuying Chen, Yuzhe Cen, Zhaofan Zou, Zhaohong Liu, Zhehao Shen, Zhenglin Du, Zhengyang Li, Zhenni Huang, Zhenwei Shao, Zhilong Song, Zhiyong Feng, Zhiyu Wang, Zhou Yu, Ziang Li, Zihan Zhai, Zijian Zhang, Ziyang Peng, Ziyun Xiao, Zongshu Li

Comments ICCV 2025 MARS2 Workshop and Challenge "Multimodal Reasoning and Slow Thinking in the Large Model Era: Towards System 2 and Beyond''

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17600 2025-09-18 cs.RO cs.AI cs.CV cs.LG

GWM: Towards Scalable Gaussian World Models for Robotic Manipulation

Guanxing Lu, Baoxiong Jia, Puhao Li, Yixin Chen, Ziwei Wang, Yansong Tang, Siyuan Huang

机构 * Tsinghua University(清华大学) State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI) School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)

Comments Published at ICCV 2025. Project page: https://gaussian-world-model.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06757 2025-09-18 cs.CV cs.GR

VOccl3D: A Video Benchmark Dataset for 3D Human Pose and Shape Estimation under real Occlusions

Yash Garg, Saketh Bachu, Arindam Dutta, Rohit Lal, Sarosij Bose, Calvin-Khang Ta, M. Salman Asif, Amit Roy-Chowdhury

机构 * University of California, Riverside(加州大学河滨分校)

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05631 2025-09-18 cs.CV

GAP: Gaussianize Any Point Clouds with Text Guidance

Weiqi Zhang, Junsheng Zhou, Haotian Geng, Wenyuan Zhang, Yu-Shen Liu

机构 * School of Software, Tsinghua University(软件学院,清华大学)

Comments ICCV 2025. Project page: https://weiqi-zhang.github.io/GAP

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13756 2025-09-18 cs.CV

PlaneRecTR++: Unified Query Learning for Joint 3D Planar Reconstruction and Pose Estimation

Jingjia Shi, Shuaifeng Zhi, Kai Xu

机构 * National University of Defense Technology(国防科技大学)

Comments To be published in IEEE T-PAMI 2025. This is the journal extension of our ICCV 2023 paper "PlaneRecTR", which expands from single view reconstruction to simultaneous multi-view reconstruction and camera pose estimation. Note that the ICCV2023 PlaneRecTR paper could be found in the previous arxiv version [v2](arXiv:2307.13756v2)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13185 2025-09-17 cs.LG cs.AI

Is Meta-Learning Out? Rethinking Unsupervised Few-Shot Classification with Limited Entropy

Yunchuan Guan, Yu Liu, Ke Zhou, Zhiqi Shen, Jenq-Neng Hwang, Serge Belongie, Lei Li

机构 * Huazhong University of Science and Technology(华中科技大学) Nanyang Technological University(南洋理工大学) University of Washington(华盛顿大学) University of Copenhagen(哥本哈根大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12894 2025-09-17 cs.CV

DialNav: Multi-turn Dialog Navigation with a Remote Guide

Leekyeung Han, Hyunji Min, Gyeom Hwangbo, Jonghyun Choi, Paul Hongsuck Seo

机构 * Korea University(韩国大学) University of Seoul(首尔大学) Seoul National University(首尔国立大学)

Comments 18 pages, 8 figures, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12512 2025-09-17 eess.IV cs.AI cs.CV

DinoAtten3D: Slice-Level Attention Aggregation of DinoV2 for 3D Brain MRI Anomaly Classification

Fazle Rafsani, Jay Shah, Catherine D. Chong, Todd J. Schwedt, Teresa Wu

机构 * Arizona State University(亚利桑那州立大学) Mayo Clinic, Arizona(梅奥诊所(亚利桑那))

Comments ACCEPTED at the ICCV 2025 Workshop on Anomaly Detection with Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12482 2025-09-17 cs.CV

Towards Foundational Models for Single-Chip Radar

Tianshu Huang, Akarsh Prabhakara, Chuhan Chen, Jay Karhade, Deva Ramanan, Matthew O'Toole, Anthony Rowe

机构 * Carnegie Mellon University(卡内基梅隆大学) Bosch Research(博世研究) University of Wisconsin–Madison(威斯康星大学麦迪逊分校)

Comments To appear in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07825 2025-09-17 cs.CV

3DSRBench: A Comprehensive 3D Spatial Reasoning Benchmark

Wufei Ma, Haoyu Chen, Guofeng Zhang, Yu-Cheng Chou, Jieneng Chen, Celso M de Melo, Alan Yuille

机构 * Johns Hopkins University(约翰霍普金斯大学) Carnegie Mellon University(卡内基梅隆大学) DEVCOM Army Research Laboratory(DEVCOM陆军研究实验室)

Comments ICCV 2025. Project page: https://3dsrbench.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19331 2025-09-17 cs.CV cs.AI cs.CL

Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabulary Segmentation

Luca Barsellotti, Lorenzo Bianchi, Nicola Messina, Fabio Carrara, Marcella Cornia, Lorenzo Baraldi, Fabrizio Falchi, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) ISTI-CNR(意大利国家研究委员会ISTI) University of Pisa(比萨大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11959 2025-09-16 cs.CV cs.RO

Learning to Generate 4D LiDAR Sequences

Ao Liang, Youquan Liu, Yu Yang, Dongyue Lu, Linfeng Li, Lingdong Kong, Huaici Zhao, Wei Tsang Ooi

机构 * NUS(国立新加坡大学) UCAS(中国科学院大学) SIA, CAS(中国科学院上海自动化研究所) FDU(福建大学) ZJU(浙江大学)

Comments Abstract Paper (Non-Archival) @ ICCV 2025 Wild3D Workshop; GitHub Repo at https://lidarcrafter.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11840 2025-09-16 cs.CV

Synthetic Captions for Open-Vocabulary Zero-Shot Segmentation

Tim Lebailly, Vijay Veerabadran, Satwik Kottur, Karl Ridgeway, Michael Louis Iuzzolino

机构 * Meta KU Leuven(鲁汶大学)

Comments ICCV 2025 CDEL Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07082 2025-09-16 cs.CV cs.AI cs.LG

On the Generalization of Representation Uncertainty in Earth Observation

Spyros Kondylatos, Nikolaos Ioannis Bountos, Dimitrios Michail, Xiao Xiang Zhu, Gustau Camps-Valls, Ioannis Papoutsis

机构 * National Observatory of Athens(雅典国家天文台) National Technical University of Athens(雅典技术大学) University of Valencia(瓦伦西亚大学) Harokopio University of Athens(雅典惠克罗波利斯大学) Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心) Archimedes/Athena RC(阿基米德/雅典RC)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08641 2025-09-16 cs.CV

3D Mesh Editing using Masked LRMs

Will Gao, Dilin Wang, Yuchen Fan, Aljaz Bozic, Tuur Stuyck, Zhengqin Li, Zhao Dong, Rakesh Ranjan, Nikolaos Sarafianos

机构 * University of Chicago(芝加哥大学) Meta Reality Labs(Meta现实实验室)

Comments ICCV 2025. Project Page: https://chocolatebiscuit.github.io/MaskedLRM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.10123 2025-09-16 cs.RO cs.CV

Learning Precise Affordances from Egocentric Videos for Robotic Manipulation

Gen Li, Nikolaos Tsagkas, Jifei Song, Ruaridh Mon-Williams, Sethu Vijayakumar, Kun Shao, Laura Sevilla-Lara

机构 * University of Edinburgh(爱丁堡大学) Huawei Noah’s Ark Lab(华为诺亚实验室)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11674 2025-09-16 cs.CV

RouteExtract: A Modular Pipeline for Extracting Routes from Paper Maps

Bjoern Kremser, Yusuke Matsui

机构 * Technical University of Munich(慕尼黑技术大学) The University of Tokyo(东京大学)

Comments Accepted to the Workshop on Graphic Design Understanding and Generation (GDUG) at ICCV 2025. 8 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11476 2025-09-16 cs.CV cs.LG

Modality-Aware Infrared and Visible Image Fusion with Target-Aware Supervision

Tianyao Sun, Dawei Xiang, Tianqi Ding, Xiang Fang, Yijiashun Qi, Zunduo Zhao

机构 * Independent researcher(独立研究者) Dept. of Computer Science Baylor University(计算机科学系 巴里尔大学) Dept. of Computer Science Engineering University of Connecticut(计算机科学工程系 佛罗里达大学) Dept. of Computer Science New York University(计算机科学系 新 york 大学)

Comments Accepted by 2025 6th International Conference on Computer Vision and Data Mining (ICCVDM 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11394 2025-09-16 cs.CV

MixANT: Observation-dependent Memory Propagation for Stochastic Dense Action Anticipation

Syed Talal Wasim, Hamid Suleman, Olga Zatsarynna, Muzammal Naseer, Juergen Gall

机构 * University of Bonn(波恩大学) Lamarr Institute of ML and AI(拉马尔人工智能与机器学习研究所) Khalifa University(卡利法大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11054 2025-09-16 cs.IT cs.CV math.IT

Rate-Distortion Limits for Multimodal Retrieval: Theory, Optimal Codes, and Finite-Sample Guarantees

Thomas Y. Chen

机构 * Department of Computer Science, Columbia University(计算机科学系,哥伦比亚大学)

Comments ICCV MRR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10620 2025-09-16 cs.CV cs.LG

Building a General SimCLR Self-Supervised Foundation Model Across Neurological Diseases to Advance 3D Brain MRI Diagnoses

Emily Kaczmarek, Justin Szeto, Brennan Nichyporuk, Tal Arbel

机构 * McGill University(麦吉尔大学) Mila - Quebec Artificial Intelligence Institute(魁北克人工智能研究所)

Comments Accepted to ICCV 2025 Workshop CVAMD

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06336 2025-09-16 cs.CV cs.AI cs.CR

Multi-View Slot Attention Using Paraphrased Texts for Face Anti-Spoofing

Jeongmin Yu, Susang Kim, Kisu Lee, Taekyoung Kwon, Won-Yong Shin, Ha Young Kim

机构 * Yonsei University(延世大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01020 2025-09-16 cs.CV

Shot-by-Shot: Film-Grammar-Aware Training-Free Audio Description Generation

Junyu Xie, Tengda Han, Max Bain, Arsha Nagrani, Eshika Khandelwal, Gül Varol, Weidi Xie, Andrew Zisserman

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组) CVIT, IIIT Hyderabad(海得拉巴印度理工学院计算机视觉研究所) LIGM, École des Ponts ParisTech(巴黎理工学院路易-狄塞尔数学与计算机科学实验室) SAI, Shanghai Jiao Tong University(上海交通大学人工智能研究所)

Comments ICCV 2025. Project Page: https://www.robots.ox.ac.uk/vgg/research/shot-by-shot/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01428 2025-09-16 cs.CV eess.IV

DLF: Extreme Image Compression with Dual-generative Latent Fusion

Naifu Xue, Zhaoyang Jia, Jiahao Li, Bin Li, Yuan Zhang, Yan Lu

机构 * Communication University of China(中国通信大学) University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13389 2025-09-16 cs.CV cs.LG

Marigold-DC: Zero-Shot Monocular Depth Completion with Guided Diffusion

Massimiliano Viola, Kevin Qu, Nando Metzger, Bingxin Ke, Alexander Becker, Konrad Schindler, Anton Obukhov

机构 * ETH Zürich(苏黎世联邦理工学院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏