arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2503.07535 2025-08-25 cs.CV

LBM: Latent Bridge Matching for Fast Image-to-Image Translation

Clément Chadebec, Onur Tasar, Sanjeev Sreetharan, Benjamin Aubin

机构 * Jasper Research(杰斯珀研究)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15774 2025-08-22 cs.CV

CineScale: Free Lunch in High-Resolution Cinematic Visual Generation

Haonan Qiu, Ning Yu, Ziqi Huang, Paul Debevec, Ziwei Liu

机构 * Nanyang Technological University(南洋理工大学) Netflix Eyeline Studios

Comments CineScale is an extended work of FreeScale (ICCV 2025). Project Page: https://eyeline-labs.github.io/CineScale/, Code Repo: https://github.com/Eyeline-Labs/CineScale

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15767 2025-08-22 cs.CV

ATLAS: Decoupling Skeletal and Shape Parameters for Expressive Parametric Human Modeling

Jinhyung Park, Javier Romero, Shunsuke Saito, Fabian Prada, Takaaki Shiratori, Yichen Xu, Federica Bogo, Shoou-I Yu, Kris Kitani, Rawal Khirodkar

机构 * Meta Carnegie Mellon University(卡内基梅隆大学)

Comments ICCV 2025; Website: https://jindapark.github.io/projects/atlas/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15752 2025-08-22 cs.HC cs.AI cs.CV

"Does the cafe entrance look accessible? Where is the door?" Towards Geospatial AI Agents for Visual Inquiries

Jon E. Froehlich, Jared Hwang, Zeyu Wang, John S. O'Meara, Xia Su, William Huang, Yang Zhang, Alex Fiannaca, Philip Nelson, Shaun Kane

机构 * University of Washington(华盛顿大学) Google Research(谷歌研究) UCLA(加州大学洛杉矶分校) Google DeepMind(谷歌DeepMind)

Comments Accepted to the ICCV'25 Workshop "Vision Foundation Models and Generative AI for Accessibility: Challenges and Opportunities"

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11988 2025-08-22 cs.CV

Exploring Spatial-Temporal Dynamics in Event-based Facial Micro-Expression Analysis

Nicolas Mastropasqua, Ignacio Bugueno-Cordova, Rodrigo Verschae, Daniel Acevedo, Pablo Negri, Maria E. Buemi

机构 * Universidad de Buenos Aires, Facultad de Ciencias Exactas y Naturales(布宜诺斯艾利斯大学,精确科学与自然学院) Institute of Engineering Sciences, Universidad de O’Higgins(工程科学研究所,奥希金斯大学) CONICET-UBA, Instituto de Ciencias de la Computacion (ICC)(CONICET-UBA,计算科学研究所) L3S Research Center, Leibniz Universität Hannover(L3S研究中心,汉诺威莱布尼茨大学)

Journal ref 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW); 2nd Workshop on Neuromorphic Vision (NeVi)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15537 2025-08-22 cs.CV

D3FNet: A Differential Attention Fusion Network for Fine-Grained Road Structure Extraction in Remote Perception Systems

Chang Liu, Yang Xu, Tamas Sziranyi

Comments 10 pages, 6 figures, International Conference on Computer Vision, ICCV 2025 (DriveX) paper id 5

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13560 2025-08-22 cs.CV

DictAS: A Framework for Class-Generalizable Few-Shot Anomaly Segmentation via Dictionary Lookup

Zhen Qu, Xian Tao, Xinyi Gong, ShiChen Qu, Xiaopei Zhang, Xingang Wang, Fei Shen, Zhengtao Zhang, Mukesh Prasad, Guiguang Ding

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Casivision Longmen Laboratory(龙门实验室) HDU UTS UCLA(加州大学洛杉矶分校) Tsinghua University(清华大学)

Comments Accepted by ICCV 2025, Project: https://github.com/xiaozhen228/DictAS

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10225 2025-08-22 cs.CV

Synthesizing Near-Boundary OOD Samples for Out-of-Distribution Detection

Jinglun Li, Kaixun Jiang, Zhaoyu Chen, Bo Lin, Yao Tang, Weifeng Ge, Wenqiang Zhang

机构 * College of Intelligent Robotics and Advanced Manufacturing, Fudan University, Shanghai(智能机器人与先进制造学院,复旦大学,上海) Shanghai Key Lab of Intelligent Information Processing, College of Computer Science and Artificial Intelligence, Fudan University, Shanghai(上海智能信息处理重点实验室,计算机科学与人工智能学院,复旦大学,上海) JIIOV Technology, Beijing(JIIOV技术,北京)

Comments Accepted by ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13787 2025-08-22 cs.CV cs.LG

Adaptive Routing of Text-to-Image Generation Requests Between Large Cloud Model and Light-Weight Edge Model

Zewei Xin, Qinya Li, Chaoyue Niu, Fan Wu, Guihai Chen

机构 * Shanghai Jiao Tong University(上海交通大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12892 2025-08-22 cs.CV

3DGS-LM: Faster Gaussian-Splatting Optimization with Levenberg-Marquardt

Lukas Höllein, Aljaž Božič, Michael Zollhöfer, Matthias Nießner

机构 * Technical University of Munich(慕尼黑技术大学) Meta

Comments Accepted to ICCV 2025. Project page: https://lukashoel.github.io/3DGS-LM, Video: https://www.youtube.com/watch?v=tDiGuGMssg8, Code: https://github.com/lukasHoel/3DGS-LM

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17840 2025-08-22 cs.AI cs.CV

Human-Object Interaction from Human-Level Instructions

Zhen Wu, Jiaman Li, Pei Xu, C. Karen Liu

机构 * Stanford University(斯坦福大学)

Comments ICCV 2025, project page: https://hoifhli.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.08207 2025-08-22 cs.CV

Translating Images to Road Network: A Sequence-to-Sequence Perspective

Jiachen Lu, Ming Nie, Bozhou Zhang, Reyuan Peng, Xinyue Cai, Hang Xu, Feng Wen, Wei Zhang, Li Zhang

机构 * School of Data Science, Fudan University(复旦大学数据科学学院) Huawei Noah’s Ark Lab(华为诺亚实验室)

Comments V1 is the ICCV 2023 conference version, and V2 is the extended version

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14588 2025-08-21 cs.CV

Controllable Latent Space Augmentation for Digital Pathology

Sofiène Boutaj, Marin Scalbert, Pierre Marza, Florent Couzinie-Devy, Maria Vakalopoulou, Stergios Christodoulidis

机构 * MICS, CentraleSupélec – Université Paris-Saclay(MICS,中央圣艾尔布兰理工大学——巴黎-萨克雷大学) Bioptimus, Inc.(Bioptimus公司) VitaDX International(VitaDX国际)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04171 2025-08-21 cs.CV

DuCos: Duality Constrained Depth Super-Resolution via Foundation Model

Zhiqiang Yan, Zhengxue Wang, Haoye Dong, Jun Li, Jian Yang, Gim Hee Lee

机构 * National University of Singapore(国立新加坡大学) Nanjing University of Science and Technology(南京理工大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14440 2025-08-21 cs.CV

MUSE: Multi-Subject Unified Synthesis via Explicit Layout Semantic Expansion

Fei Peng, Junqiang Wu, Yan Li, Tingting Gao, Di Zhang, Huiyuan Fu

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Kuaishou Technology(快手科技)

Comments This paper is accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14358 2025-08-21 cs.CV cs.AI cs.RO

Learning Point Cloud Representations with Pose Continuity for Depth-Based Category-Level 6D Object Pose Estimation

Zhujun Li, Shuo Zhang, Ioannis Stamos

机构 * Graduate Center, CUNY(CUNY研究生中心) Hunter College, CUNY(CUNY霍普金斯学院) Weill Cornell Medicine(韦尔医学院)

Comments Accepted by ICCV 2025 Workshop on Recovering 6D Object Pose (R6D)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12911 2025-08-21 cs.RO cs.LG

LaViPlan : Language-Guided Visual Path Planning with RLVR

Hayeon Oh

机构 * Electronics and Telecommunications Research Institute(电子通信研究院)

Comments Accepted to the 2nd ICCV 2025 Workshop on the Challenge of Out-of-Label Hazards in Autonomous Driving (13 pages, 6 figures)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.05311 2025-08-21 cs.CV

MMAD: Multi-label Micro-Action Detection in Videos

Kun Li, Pengyu Liu, Dan Guo, Fei Wang, Zhiliang Wu, Hehe Fan, Meng Wang

机构 * School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院) ReLER, CCAI, Zhejiang University(浙江大学ReLER、CCAI)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14041 2025-08-20 cs.CV

LongSplat: Robust Unposed 3D Gaussian Splatting for Casual Long Videos

Chin-Yang Lin, Cheng Sun, Fu-En Yang, Min-Hung Chen, Yen-Yu Lin, Yu-Lun Liu

机构 * National Yang Ming Chiao Tung University

Comments ICCV 2025. Project page: https://linjohnss.github.io/longsplat/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14039 2025-08-20 cs.CV

Beyond Simple Edits: Composed Video Retrieval with Dense Modifications

Omkar Thawakar, Dmitry Demidov, Ritesh Thawkar, Rao Muhammad Anwer, Mubarak Shah, Fahad Shahbaz Khan, Salman Khan

机构 * Mohamed bin Zayed University of AI(马尔代夫比兹赞大学人工智能学院) University of Central Florida(中央佛罗里达大学) Linköping University(林霍尔姆大学) Australian National University(澳大利亚国立大学)

Comments Accepted to ICCV-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14015 2025-08-20 cs.CV

Backdooring Self-Supervised Contrastive Learning by Noisy Alignment

Tuo Chen, Jie Gui, Minjing Dong, Ju Jia, Lanting Fang, Jian Liu

机构 * Southeast University(东南大学) Purple Mountain Laboratories(紫金山实验室) Ant Group(蚂蚁集团) City University of Hong Kong(香港城市大学) Beijing Institute of Technology(北京理工大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13880 2025-08-20 cs.CV

In-hoc Concept Representations to Regularise Deep Learning in Medical Imaging

Valentina Corbetta, Floris Six Dijkstra, Regina Beets-Tan, Hoel Kervadec, Kristoffer Wickstrøm, Wilson Silva

机构 * The Netherlands Cancer Institute(荷兰癌症研究所) Utrecht University(乌得勒支大学) Maastricht University(马斯特里赫特大学) University of Amsterdam(阿姆斯特丹大学) Amsterdam UMC(阿姆斯特丹大学医学中心) UiT The Arctic University of Norway(挪威北莫斯堡大学)

Comments 13 pages, 13 figures, 2 tables, accepted at PHAROS-AFE-AIMI Workshop in conjunction with the International Conference on Computer Vision (ICCV), 2025. This is the submitted manuscript with added link to the github repo, funding acknowledgments and author names and affiliations, and a correction to numbers in Table 1. Final version not published yet

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13564 2025-08-20 cs.CV cs.AI cs.LG cs.RO

The 9th AI City Challenge

Zheng Tang, Shuo Wang, David C. Anastasiu, Ming-Ching Chang, Anuj Sharma, Quan Kong, Norimasa Kobori, Munkhjargal Gochoo, Ganzorig Batnasan, Munkh-Erdene Otgonbold, Fady Alnajjar, Jun-Wei Hsieh, Tomasz Kornuta, Xiaolong Li, Yilin Zhao, Han Zhang, Subhashree Radhakrishnan, Arihant Jain, Ratnesh Kumar, Vidya N. Murali, Yuxing Wang, Sameer Satish Pusegaonkar, Yizhou Wang, Sujit Biswas, Xunlei Wu, Zhedong Zheng, Pranamesh Chakraborty, Rama Chellappa

机构 * NVIDIA Corporation(NVIDIA公司) Santa Clara University(圣克拉拉大学) University at Albany, SUNY(纽约州立大学阿尔巴尼分校) Iowa State University(爱荷华州立大学) Woven by Toyota, Japan(日本丰田公司) United Arab Emirates University(阿拉伯联合酋长国大学) National Yang-Ming Chiao-Tung University(国家阳明交通大学) University of Macau(澳门大学) Indian Institute of Technology Kanpur(印度理工学院坎普尔分校) Johns Hopkins University(约翰霍普金斯大学)

Comments Summary of the 9th AI City Challenge Workshop in conjunction with ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13503 2025-08-20 cs.CV eess.IV

AdaptiveAE: An Adaptive Exposure Strategy for HDR Capturing in Dynamic Scenes

Tianyi Xu, Fan Zhang, Boxin Shi, Tianfan Xue, Yujin Wang

机构 * Shanghai AI Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学) State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(北京大学多媒体信息处理国家重点实验室,计算机科学学院) National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(北京大学视觉技术国家工程研究中心,计算机科学学院)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13470 2025-08-20 cs.CV cs.AI

STER-VLM: Spatio-Temporal With Enhanced Reference Vision-Language Models

Tinh-Anh Nguyen-Nhu, Triet Dao Hoang Minh, Dat To-Thanh, Phuc Le-Gia, Tuan Vo-Lan, Tien-Huy Nguyen

机构 * Ho Chi Minh University of Technology(胡志明理工大学) Vietnamese-German University(越德大学) Ho Chi Minh University of Science(胡志明理工大学) University of Information Technology(信息科技大学) Vietnam National University, Ho Chi Minh city(越南国家大学,胡志明市)

Comments ICCV Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12615 2025-08-20 cs.CV

WIPES: Wavelet-based Visual Primitives

Wenhao Zhang, Hao Zhu, Delong Wu, Di Kang, Linchao Bao, Xun Cao, Zhan Ma

机构 * Nanjing University(南京大学) Tencent(腾讯)

Comments IEEE/CVF International Conference on Computer Vision 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01996 2025-08-20 cs.LG cs.CV

Always Skip Attention

Yiping Ji, Hemanth Saratchandran, Peyman Moghadam, Simon Lucey

机构 * Adelaide University(阿德莱德大学) CSIRO(澳大利亚联邦科学与工业研究组织) Queensland University of Technology(昆士兰理工大学)

Comments This work has just been accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10486 2025-08-20 cs.CV

DNF-Avatar: Distilling Neural Fields for Real-time Animatable Avatar Relighting

Zeren Jiang, Shaofei Wang, Siyu Tang

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组) ETH Zürich(苏黎世联邦理工学院)

Comments 17 pages, 9 figures, ICCV 2025 Findings Oral, Project pages: https://jzr99.github.io/DNF-Avatar/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07961 2025-08-20 cs.CV

Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction

Zeren Jiang, Chuanxia Zheng, Iro Laina, Diane Larlus, Andrea Vedaldi

机构 * Visual Geometry Group, University of Oxford(视觉几何组,牛津大学) Naver Labs Europe(Naver欧洲实验室)

Comments 17 pages, 6 figures, ICCV 2025 Highlight, Project page: https://geo4d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22884 2025-08-20 cs.CV

AutoComPose: Automatic Generation of Pose Transition Descriptions for Composed Pose Retrieval Using Multimodal LLMs

Yi-Ting Shen, Sungmin Eum, Doheon Lee, Rohit Shete, Chiao-Yi Wang, Heesung Kwon, Shuvra S. Bhattacharyya

机构 * University of Maryland, College Park(马里兰大学学院市分校) DEVCOM Army Research Laboratory(陆军研究实验室)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏