arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2503.08829 2025-08-27 cs.LG cs.CR

Seal Your Backdoor with Variational Defense

Ivan Sabolić, Matej Grcić, Siniša Šegvić

机构 * University of Zagreb, Faculty of Electrical Engineering and Computing(Zagreb大学电气工程与计算学院)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14137 2025-08-27 cs.CV cs.CL

VAGUE: Visual Contexts Clarify Ambiguous Expressions

Heejeong Nam, Jinwoo Ahn, Keummin Ka, Jiwan Chung, Youngjae Yu

机构 * Brown University(布朗大学) UC Berkeley(加州大学伯克利分校) Yonsei University(延世大学)

Comments ICCV 2025, 32 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13195 2025-08-26 cs.CV

CoMPaSS: Enhancing Spatial Understanding in Text-to-Image Diffusion Models

Gaoyang Zhang, Bingtao Fu, Qingnan Fan, Qi Zhang, Runxing Liu, Hong Gu, Huaqi Zhang, Xinguo Liu

机构 * Zhejiang University(浙江大学) vivo Ant Group(蚂蚁集团)

Comments 21 pages, 12 figures. Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17844 2025-08-26 cs.CV cs.LG

Diffusion-Based Data Augmentation for Medical Image Segmentation

Maham Nazir, Muhammad Aqeel, Francesco Setti

机构 * School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院) Dept. of Engineering for Innovation Medicine, University of Verona(威尼斯大学创新医学工程系)

Comments Accepted to CVAMD Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17827 2025-08-26 cs.CV cs.LG

A Contrastive Learning-Guided Confident Meta-learning for Zero Shot Anomaly Detection

Muhammad Aqeel, Danijel Skocaj, Marco Cristani, Francesco Setti

机构 * Dept. of Engineering for Innovation Medicine, University of Verona(创新医学工程系,威尼斯大学) Faculty of Computer and Information Science, University of Ljubljana(计算机与信息科学系,卢布尔雅那大学) Qualyco S.r.l.(Qualyco公司)

Comments Accepted to VISION Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17817 2025-08-26 cs.CV

TemCoCo: Temporally Consistent Multi-modal Video Fusion with Visual-Semantic Collaboration

Meiqi Gong, Hao Zhang, Xunpeng Yi, Linfeng Tang, Jiayi Ma

机构 * Electronic Information School, Wuhan University, Wuhan 430072, China(武汉大学电子信息学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17789 2025-08-26 cs.CV cs.LG

Robust Anomaly Detection in Industrial Environments via Meta-Learning

Muhammad Aqeel, Shakiba Sharifi, Marco Cristani, Francesco Setti

机构 * Dept. of Engineering for Innovation Medicine, University of Verona(创新医学工程系,威尼斯大学) Qualyco S.r.l.(Qualyco公司)

Comments Accepted to VISION Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17636 2025-08-26 cs.CV cs.AI

Few-Shot Pattern Detection via Template Matching and Regression

Eunchan Jo, Dahyun Kang, Sanghyun Kim, Yunseon Choi, Minsu Cho

机构 * Pohang University of Science and Technology (POSTECH)(浦项科学技术大学)

Comments Accepted to ICCV 2025 (highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17595 2025-08-26 cs.CV

TinyGiantVLM: A Lightweight Vision-Language Architecture for Spatial Reasoning under Resource Constraints

Vinh-Thuan Ly, Hoang M. Truong, Xuan-Huong Nguyen

机构 * University of Science, VNU-HCM(越南胡志明市国家大学) Vietnam National University(越南国家大学)

Comments Accepted for presentation at the IEEE/CVF International Conference on Computer Vision (ICCV) Workshops, 2025

Journal ref IEEE/CVF International Conference on Computer Vision (ICCV) Workshops, Hawaii, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17342 2025-08-26 cs.GR cs.CV cs.MM cs.SD

DanceEditor: Towards Iterative Editable Music-driven Dance Generation with Open-Vocabulary Descriptions

Hengyuan Zhang, Zhe Li, Xingqun Qi, Mengze Li, Muyi Sun, Man Zhang, Sirui Han

机构 * Peking University(北京大学) The Hong Kong University of Science and Technology(香港科技大学) Beijing University of Posts and Telecommunications(北京邮电大学)

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16212 2025-08-26 cs.CV cs.AI cs.LG

OmniCache: A Trajectory-Oriented Global Perspective on Training-Free Cache Reuse for Diffusion Transformer Models

Huanpeng Chu, Wei Wu, Guanyu Fen, Yutao Zhang

机构 * Zhipu AI(智谱AI)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02072 2025-08-26 eess.IV

HyTIP: Hybrid Temporal Information Propagation for Masked Conditional Residual Video Coding

Yi-Hsin Chen, Yi-Chen Yao, Kuan-Wei Ho, Chun-Hung Wu, Huu-Tai Phung, Martin Benjak, Jörn Ostermann, Wen-Hsiao Peng

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20562 2025-08-26 cs.CV cs.AI

MemoryTalker: Personalized Speech-Driven 3D Facial Animation via Audio-Guided Stylization

Hyung Kyu Kim, Sangmin Lee, Hak Gu Kim

机构 * Chung-Ang University(Chung-Ang 大学) Korea University(韩国大学)

Comments Accepted in ICCV 2025; Project Page: https://cau-irislab.github.io/ICCV25-MemoryTalker/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17589 2025-08-26 cs.AI

Taming the Untamed: Graph-Based Knowledge Retrieval and Reasoning for MLLMs to Conquer the Unknown

Bowen Wang, Zhouqiang Jiang, Yasuaki Susumu, Shotaro Miwa, Tianwei Chen, Yuta Nakashima

机构 * Osaka University(大阪大学) Mitsubishi Electric Corp.(三菱电机公司)

Comments Aligned with ICCV 2025 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16827 2025-08-26 cs.GR cs.CV cs.LG

Beyond Blur: A Fluid Perspective on Generative Diffusion Models

Grzegorz Gruszczynski, Jakub Meixner, Michal Jan Wlodarczyk, Przemyslaw Musialski

Comments ICCV 2025 main conference, 8 pages paper, 20 pages appendix, 24 figures, supplementary pseudocode in appendix, https://iccv.thecvf.com/virtual/2025/poster/1176

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15123 2025-08-26 cs.CV cs.AI

Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual Grounding

Ta Duc Huy, Duy Anh Huynh, Yutong Xie, Yuankai Qi, Qi Chen, Phi Le Nguyen, Sen Kim Tran, Son Lam Phung, Anton van den Hengel, Zhibin Liao, Minh-Son To, Johan W. Verjans, Vu Minh Hieu Phan

机构 * Australian Institute for Machine Learning, University of Adelaide(澳大利亚机器学习研究所,阿德莱德大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) Macquarie University(麦考瑞大学) Hanoi University of Science and Technology(河内科学技术大学) University of Wollongong(沃林根大学) Flinders University(弗林德斯大学)

Comments Accepted at ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08736 2025-08-26 cs.CV

GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation

Tianwei Xiong, Jun Hao Liew, Zilong Huang, Jiashi Feng, Xihui Liu

机构 * The University of Hong Kong(香港大学) ByteDance Seed Project(字节跳动种子项目)

Comments ICCV 2025. Project page: https://silentview.github.io/GigaTok

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17288 2025-08-26 cs.CV

GaussianFlowOcc: Sparse and Weakly Supervised Occupancy Estimation using Gaussian Splatting and Temporal Flow

Simon Boeder, Fabian Gigengack, Benjamin Risse

机构 * Robert Bosch GmbH(罗伯特·博世有限公司) University of Münster(穆斯尔大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13087 2025-08-26 cs.CV cs.LG

Orchid: Image Latent Diffusion for Joint Appearance and Geometry Generation

Akshay Krishnan, Xinchen Yan, Vincent Casser, Abhijit Kundu

机构 * Google DeepMind(谷歌DeepMind) Georgia Institute of Technology(佐治亚理工学院) Waymo

Comments Accepted to ICCV 2025. Project webpage: https://orchid3d.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04715 2025-08-26 cs.CV

Addressing Text Embedding Leakage in Diffusion-based Image Editing

Sunung Mun, Jinhwan Nam, Sunghyun Cho, Jungseul Ok

机构 * Graduate School of AI, POSTECH(POSTECH人工智能研究生院) Dept. of CSE, POSTECH(POSTECH计算机科学与工程系)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16789 2025-08-26 cs.CV cs.CL

Leveraging the Power of MLLMs for Gloss-Free Sign Language Translation

Jungeun Kim, Hyeongwoo Jeon, Jongseong Bae, Ha Young Kim

机构 * Department of Artificial Intelligence, Yonsei University(人工智能系,延世大学) Graduate School of Information, Yonsei University(信息研究生院,延世大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16847 2025-08-26 cs.CV cs.AI

TokenUnify: Scaling Up Autoregressive Pretraining for Neuron Segmentation

Yinda Chen, Haoyuan Shi, Xiaoyu Liu, Te Shi, Ruobing Zhang, Dong Liu, Zhiwei Xiong, Feng Wu

机构 * MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知大学科学与技术大学实验室) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究所) Institute for Brain and Intelligence, Fudan University(脑与智能研究所)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16911 2025-08-26 cs.GR cs.CV cs.MM cs.SD

MDD: A Dataset for Text-and-Music Conditioned Duet Dance Generation

Prerit Gupta, Jason Alexander Fotso-Puepi, Zhengyuan Li, Jay Mehta, Aniket Bera

机构 * Purdue University(普渡大学)

Comments Accepted at ICCV 2025. Project page: https://gprerit96.github.io/mdd-page

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16830 2025-08-26 cs.CV eess.IV

AIM 2025 Low-light RAW Video Denoising Challenge: Dataset, Methods and Results

Alexander Yakovenko, George Chakvetadze, Ilya Khrapov, Maksim Zhelezov, Dmitry Vatolin, Radu Timofte, Youngjin Oh, Junhyeong Kwon, Junyoung Park, Nam Ik Cho, Senyan Xu, Ruixuan Jiang, Long Peng, Xueyang Fu, Zheng-Jun Zha, Xiaoping Peng, Hansen Feng, Zhanyi Tie, Ziming Xia, Lizhi Wang

机构 * AIM 2025 Low-light RAW Video Denoising Challenge(AIM 2025 低光照RAW视频去噪挑战)

Comments Challenge report from Advances in Image Manipulation workshop held at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16762 2025-08-26 cs.CL cs.CY

Toward Socially Aware Vision-Language Models: Evaluating Cultural Competence Through Multimodal Story Generation

Arka Mukherjee, Shreya Ghosh

机构 * KIIT Deemed University(KIIT大学) Indian Institute of Technology (IIT) Bhubaneswar(印度理工学院(Bhubaneswar分校))

Comments Accepted at ASI @ ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16121 2025-08-25 eess.IV cs.CV

Lightweight and Fast Real-time Image Enhancement via Decomposition of the Spatial-aware Lookup Tables

Wontae Kim, Keuntek Lee, Nam Ik Cho

机构 * IPAI, Seoul National University, Seoul, Korea(IPAI,首尔国立大学,首尔,韩国) LG Electronics, Seoul, Korea(LG电子,首尔,韩国) Department of ECE, INMC, Seoul National University, Seoul, Korea(电子工程系,INMC,首尔国立大学,首尔,韩国)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13957 2025-08-25 cs.CV

ViT-FIQA: Assessing Face Image Quality using Vision Transformers

Andrea Atzori, Fadi Boutros, Naser Damer

机构 * Fraunhofer Institute for Computer Graphics Research IGD(弗劳恩霍夫计算机图形研究 institutes IGD) Department of Computer Science, TU Darmstadt(图腾大学计算机科学系)

Comments Accepted at the IEEE/CVF International Conference on Computer Vision Workshops 2025 (ICCVW 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01225 2025-08-25 cs.CV cs.AI

Multi-Cache Enhanced Prototype Learning for Test-Time Generalization of Vision-Language Models

Xinyu Chen, Haotian Zhai, Can Zhang, Xiupeng Shi, Ruirui Li

机构 * Shanghai University(上海大学) Beijing University of Chemical Technology(北京化工大学) University of Minnesota(明尼苏达大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16779 2025-08-25 eess.IV cs.CV

Improving U-Net Confidence on TEM Image Data with L2-Regularization, Transfer Learning, and Deep Fine-Tuning

Aiden Ochoa, Xinyuan Xu, Xing Wang

机构 * Department of Nuclear Engineering, Penn State University(核工程系,宾夕法尼亚州立大学)

Comments Accepted into the ICCV 2025 CV4MS Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11652 2025-08-25 cs.CV

Bring Your Rear Cameras for Egocentric 3D Human Pose Estimation

Hiroyasu Akada, Jian Wang, Vladislav Golyanik, Christian Theobalt

机构 * Max Planck Institute for Informatics(马克斯·普朗克信息研究所)

Comments Project page: https://4dqv.mpi-inf.mpg.de/EgoRear/

Journal ref International Conference on Computer Vision 2025 (ICCV 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏