arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2503.10636 2025-08-06 cs.LG cs.CV

The Curse of Conditions: Analyzing and Improving Optimal Transport for Conditional Flow-Based Generation

Ho Kei Cheng, Alexander Schwing

Comments ICCV 2025. Project page: https://hkchengrex.github.io/C2OT

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14432 2025-08-06 cs.CV eess.IV

IntroStyle: Training-Free Introspective Style Attribution using Diffusion Features

Anand Kumar, Jiteng Mu, Nuno Vasconcelos

机构 * University of California, San Diego(加州大学圣地亚哥分校)

Comments 17 pages, 16 figures

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02890 2025-08-06 cs.CV

EvRT-DETR: Latent Space Adaptation of Image Detectors for Event-based Vision

Dmitrii Torbunov, Yihui Ren, Animesh Ghose, Odera Dim, Yonggang Cui

机构 * Brookhaven National Laboratory(布鲁克海文国家实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.11548 2025-08-06 cs.CV

Learning Interpretable Queries for Explainable Image Classification with Information Pursuit

Stefan Kolek, Aditya Chattopadhyay, Kwan Ho Ryan Chan, Hector Andrade-Loarca, Gitta Kutyniok, Réne Vidal

机构 * LMU Munich(慕尼黑大学) AWS AI Labs(AWS人工智能实验室) University of Pennsylvania(宾夕法尼亚大学) Technical University of Munich(慕尼黑技术大学) Munich Center of Machine Learning(慕尼黑机器学习中心)

Comments Published at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02645 2025-08-05 cs.CV

Evaluating Variance in Visual Question Answering Benchmarks

Nikitha SR

机构 * Media and Data Science Research Lab, Adobe(媒体与数据科学研究实验室,Adobe)

Comments Accepted in ICCV 2025 Workshop on What's Next in Multimodal Foundational Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02339 2025-08-05 cs.CV cs.RO

Correspondence-Free Fast and Robust Spherical Point Pattern Registration

Anik Sarker, Alan T. Asbeck

机构 * Dept. of Mechanical Engineering, Virginia Tech(机械工程系,弗吉尼亚理工学院)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02323 2025-08-05 cs.CV

Dream-to-Recon: Monocular 3D Reconstruction with Diffusion-Depth Distillation from Single Images

Philipp Wulff, Felix Wimbauer, Dominik Muhle, Daniel Cremers

机构 * Technical University of Munich(慕尼黑技术大学) MCML SE3 Labs(SE3实验室)

Comments ICCV 2025. Website: https://philippwulff.github.io/dream-to-recon

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02288 2025-08-05 cs.CV

Unleashing the Temporal Potential of Stereo Event Cameras for Continuous-Time 3D Object Detection

Jae-Young Kang, Hoonhee Cho, Kuk-Jin Yoon

机构 * KAIST(韩国科学技术院)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00443 2025-08-05 cs.CV

SDMatte: Grafting Diffusion Models for Interactive Matting

Longfei Huang, Yu Liang, Hao Zhang, Jinwei Chen, Wei Dong, Lunde Chen, Wanyu Liu, Bo Li, Peng-Tao Jiang

机构 * Shanghai University(上海大学) vivo Mobile Communication Co., Ltd.(vivo移动通信有限公司)

Comments Accepted at ICCV 2025, 11 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12049 2025-08-05 cs.CV

TACO: Taming Diffusion for in-the-wild Video Amodal Completion

Ruijie Lu, Yixin Chen, Yu Liu, Jiaxiang Tang, Junfeng Ni, Diwen Wan, Gang Zeng, Siyuan Huang

机构 * State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室) State Key Laboratory of General Artificial Intelligence, BIGAI(BIGAI 通用人工智能国家重点实验室) Tsinghua University(清华大学)

Comments Accepted by ICCV 2025.Project page: https://jason-aplp.github.io/TACO

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06453 2025-08-05 cs.CR

Efficient Input-level Backdoor Defense on Text-to-Image Synthesis via Neuron Activation Variation

Shengfang Zhai, Jiajun Li, Yue Liu, Huanran Chen, Zhihua Tian, Wenjie Qu, Qingni Shen, Ruoxi Jia, Yinpeng Dong, Jiaheng Zhang

Comments 20 pages. ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17446 2025-08-05 eess.IV cs.CV

Comparing ImageNet Pre-training with Digital Pathology Foundation Models for Whole Slide Image-Based Survival Analysis

Kleanthis Marios Papadopoulos, Tania Stathaki

机构 * Imperial College London, Department of Electrical and Electronic Engineering, London, UK(帝国理工学院伦敦分校,电子与电气工程系) Department of Pathology, Institut de Pathologie Multisite, Groupement Hospitalier Sud, Lyon University Hospital, Pierre-Bénite, France(病理学系,多站点病理研究所,南部医院集团,里昂大学医院,皮埃尔-贝内蒂,法国) University of Lyon, Université Claude Bernard Lyon 1, Lyon, France(里昂大学, Claude Bernard 里昂第一大学,里昂,法国)

Comments Accepted (Oral) at the 6th International Conference on Computer Vision and Information Technology (CVIT 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02190 2025-08-05 cs.RO cs.AI

FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation

Cui Miao, Tao Chang, Meihan Wu, Hongbin Xu, Chun Li, Ming Li, Xiaodong Wang

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02186 2025-08-05 cs.CV

Failure Cases Are Better Learned But Boundary Says Sorry: Facilitating Smooth Perception Change for Accuracy-Robustness Trade-Off in Adversarial Training

Yanyun Wang, Li Liu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

Comments 2025 IEEE/CVF International Conference on Computer Vision (ICCV'25)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02157 2025-08-05 cs.CV

Unified Category-Level Object Detection and Pose Estimation from RGB Images using 3D Prototypes

Tom Fischer, Xiaojie Zhang, Eddy Ilg

机构 * Saarland University(萨尔兰大学) University of Technology Nuremberg(纽伦堡技术大学)

Comments Published at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02134 2025-08-05 cs.CV

Free-MoRef: Instantly Multiplexing Context Perception Capabilities of Video-MLLMs within Single Inference

Kuo Wang, Quanlong Zheng, Junlin Xie, Yanhao Zhang, Jinguo Luo, Haonan Lu, Liang Lin, Fan Zhou, Guanbin Li

机构 * Sun Yat-sen University(中山大学) Peng Cheng Laboratory(鹏城实验室) OPPO AI Center(OPPO人工智能中心) Research Institute(研究 institute) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Harbin Institute of Technology(哈尔滨工业大学) Shenzhen Key Laboratory of Digital Living Network and Content Service(深圳数字生活网络与内容服务重点实验室) Guangdong Key Laboratory of Big Data Analysis and Processing(广东省大数据分析与处理重点实验室)

Comments published in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02106 2025-08-05 cs.CV cs.RO

Towards Immersive Human-X Interaction: A Real-Time Framework for Physically Plausible Motion Synthesis

Kaiyang Ji, Ye Shi, Zichen Jin, Kangyi Chen, Lan Xu, Yuexin Ma, Jingyi Yu, Jingya Wang

机构 * ShanghaiTech University(上海科技大学) Shanghai Engineering Research Center of Intelligent Vision and Imaging(上海智能视觉与成像工程研究中心)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02047 2025-08-05 cs.CV

Mapillary Vistas Validation for Fine-Grained Traffic Signs: A Benchmark Revealing Vision-Language Model Limitations

Sparsh Garg, Abhishek Aich

机构 * NEC Laboratories, America(美国 NEC 实验室)

Comments Accepted to ICCV 2025 Workshop (4th DataCV Workshop and Challenge)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01984 2025-08-05 cs.CV

IMoRe: Implicit Program-Guided Reasoning for Human Motion Q&A

Chen Li, Chinthani Sugandhika, Yeo Keat Ee, Eric Peh, Hao Zhang, Hong Yang, Deepu Rajan, Basura Fernando

机构 * Institute of High-Performance Computing, Agency for Science, Technology and Research, Singapore(高性能计算研究所,科技研究局,新加坡) Centre for Frontier AI Research, Agency for Science, Technology and Research, Singapore(前沿人工智能研究中心,科技研究局,新加坡) College of Computing and Data Science, Nanyang Technological University, Singapore(计算与数据科学学院,南洋理工大学,新加坡)

Comments *Equal contribution. Accepted by the International Conference on Computer Vision (ICCV 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01921 2025-08-05 cs.CV

InspectVLM: Unified in Theory, Unreliable in Practice

Conor Wallace, Isaac Corley, Jonathan Lwowski

机构 * Zeitview

Comments Accepted to 2025 ICCV VISION Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01749 2025-08-05 cs.CV cs.AI

Improving Noise Efficiency in Privacy-preserving Dataset Distillation

Runkai Zheng, Vishnu Asutosh Dasu, Yinong Oliver Wang, Haohan Wang, Fernando De la Torre

机构 * Carnegie Mellon University(卡内基梅隆大学) Pennsylvania State University(宾夕法尼亚州立大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01728 2025-08-05 cs.CV cs.AI

Granular Concept Circuits: Toward a Fine-Grained Circuit Discovery for Concept Representations

Dahee Kwon, Sehyun Lee, Jaesik Choi

机构 * KAIST AI(韩国科学技术院人工智能研究所)

Comments ICCV 2025 accepted paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01639 2025-08-05 cs.CV

Glass Surface Segmentation with an RGB-D Camera via Weighted Feature Fusion for Service Robots

Henghong Lin, Zihan Zhu, Tao Wang, Anastasia Ioannou, Yuanshui Huang

机构 * Fujian Provincial Key Laboratory of Information Processing and Intelligent Control(福建省信息处理与智能控制重点实验室) Minjiang University(闽江学院) European University Cyprus(塞浦路斯欧洲大学) Fujian Hantewin Intelligent Technology Co., Ltd.(福建省翰文智能科技有限公司)

Comments Paper accepted by 6th International Conference on Computer Vision, Image and Deep Learning (CVIDL 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01251 2025-08-05 cs.LG

Soft Separation and Distillation: Toward Global Uniformity in Federated Unsupervised Learning

Hung-Chieh Fang, Hsuan-Tien Lin, Irwin King, Yifei Zhang

机构 * National Taiwan University(国立台湾大学) The Chinese University of Hong Kong(香港中文大学)

Comments Published at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01152 2025-08-05 cs.CV

LawDIS: Language-Window-based Controllable Dichotomous Image Segmentation

Xinyu Yan, Meijun Sun, Ge-Peng Ji, Fahad Shahbaz Khan, Salman Khan, Deng-Ping Fan

机构 * Tianjin University(天津大学) Tianjin Key Laboratory of Machine Learning(天津机器学习重点实验室) Australian National University(澳大利亚国立大学) Nankai Institute of Advanced Research (SHENZHEN FUTIAN)(南开先进研究院(深圳福田)) Nankai University(南开大学) MBZUAI

Comments 17 pages, 10 figures, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01112 2025-08-05 cs.CV

MASIV: Toward Material-Agnostic System Identification from Videos

Yizhou Zhao, Haoyu Chen, Chunjiang Liu, Zhenyang Li, Charles Herrmann, Junhwa Hur, Yinxiao Li, Ming-Hsuan Yang, Bhiksha Raj, Min Xu

机构 * Carnegie Mellon University(卡内基梅隆大学) University of Alabama at Birmingham(阿拉巴马大学伯明翰分校) Google(谷歌) UC Merced(加州大学默塞德分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01098 2025-08-05 cs.CV

Trans-Adapter: A Plug-and-Play Framework for Transparent Image Inpainting

Yuekun Dai, Haitian Li, Shangchen Zhou, Chen Change Loy

机构 * S-Lab, Nanyang Technological University(南洋理工大学S实验室)

Comments accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01087 2025-08-05 cs.CV

COSTARR: Consolidated Open Set Technique with Attenuation for Robust Recognition

Ryan Rabinowitz, Steve Cruz, Walter Scheirer, Terrance E. Boult

机构 * University of Colorado Colorado Springs(科罗拉多州立大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01074 2025-08-05 cs.CV cs.CR

Evading Data Provenance in Deep Neural Networks

Hongyu Zhu, Sichu Liang, Wenwen Wang, Zhuomeng Zhang, Fangqi Li, Shi-Lin Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Southeast University(东南大学) Carnegie Mellon University(卡内基梅隆大学)

Comments ICCV 2025 Highlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01008 2025-08-05 cs.CV

ROVI: A VLM-LLM Re-Captioned Dataset for Open-Vocabulary Instance-Grounded Text-to-Image Generation

Cihang Peng, Qiming Hou, Zhong Ren, Kun Zhou

机构 * State Key Lab of CAD&CG(计算机辅助设计与图形学国家重点实验室)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏