arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4772
2507.06510 2025-07-10 cs.CV

Bilateral Collaboration with Large Vision-Language Models for Open Vocabulary Human-Object Interaction Detection

Yupeng Hu, Changxing Ding, Chang Sun, Shaoli Huang, Xiangmin Xu

机构 * South China University of Technology(南方科技大学) Tencent AI Lab(腾讯人工智能实验室) Foshan University(佛山大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23581 2025-07-10 cs.CV cs.AI cs.LG

PBCAT: Patch-based composite adversarial training against physically realizable attacks on object detection

Xiao Li, Yiming Zhu, Yifan Huang, Wei Zhang, Yingzhe He, Jie Shi, Xiaolin Hu

机构 * Department of Computer Science and Technology, BNRist, IDG/McGovern Institute for Brain Research, THBI, Tsinghua University(清华大学计算机科学与技术系) University of Science and Technology Beijing(北京科技大学) Huawei Technologies(华为技术有限公司) Chinese Institute for Brain Research (CIBR)(中国脑科学研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22637 2025-07-10 cs.CV

CaO$_2$: Rectifying Inconsistencies in Diffusion-Based Dataset Distillation

Haoxuan Wang, Zhenghao Zhao, Junyi Wu, Yuzhang Shang, Gaowen Liu, Yan Yan

机构 * University of Illinois Chicago(伊利诺伊大学香槟分校) University of Central Florida(中央佛罗里达大学) Cisco Research(思科研究)

Comments ICCV 2025. Code is available at https://github.com/hatchetProject/CaO2

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02195 2025-07-10 cs.CV

HyperGCT: A Dynamic Hyper-GNN-Learned Geometric Constraint for 3D Registration

Xiyu Zhang, Jiayi Ma, Jianwei Guo, Wei Hu, Zhaoshuai Qi, Fei Hui, Jiaqi Yang, Yanning Zhang

机构 * Northwestern Polytechnical University(北华理工大学) Wuhan University(武汉大学) Chinese Academy of Sciences(中国科学院) Peking University(北京大学) Chang’an University(长安大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09572 2025-07-10 cs.CV

Dynamic Reconstruction of Hand-Object Interaction with Distributed Force-aware Contact Representation

Zhenjun Yu, Wenqiang Xu, Pengfei Xie, Yutong Li, Brian W. Anthony, Zhuorui Zhang, Cewu Lu

机构 * Shanghai Jiao Tong University(上海交通大学) Massachusetts Institute of Technology(麻省理工学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04709 2025-07-10 cs.CV

TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation

Wenhao Wang, Yi Yang

机构 * University of Technology Sydney(悉尼技术大学) Zhejiang University(浙江大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.06079 2025-07-10 cs.CV

Towards Adversarial Robustness via Debiased High-Confidence Logit Alignment

Kejia Zhang, Juanjuan Weng, Shaozi Li, Zhiming Luo

机构 * Department of Artificial Intelligence, Xiamen University(厦门大学人工智能学院) College of Information Science and Technology, Jinan University(济南大学信息科学与技术学院) Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(教育部多媒体可信感知与高效计算重点实验室,厦门大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17399 2025-07-10 cs.CV eess.IV

Self-Calibrated Variance-Stabilizing Transformations for Real-World Image Denoising

Sébastien Herbreteau, Michael Unser

机构 * CIBM Center for Biomedical Imaging(生物医学成像中心) Biomedical Imaging Group(生物医学成像组) Univ Rennes, Ensai, CNRS, CREST—UMR 9194(里尔大学、Ensai、CNRS、CREST—UMR 9194)

Comments Accepted at IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03010 2025-07-10 cs.CV

CAVIS: Context-Aware Video Instance Segmentation

Seunghun Lee, Jiwan Seo, Kiljoon Han, Minwoo Choi, Sunghoon Im

机构 * DGIST

Comments ICCV 2025. Code: https://github.com/Seung-Hun-Lee/CAVIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14030 2025-07-10 cs.CV cs.CL

Refining Skewed Perceptions in Vision-Language Contrastive Models through Visual Representations

Haocheng Dai, Sarang Joshi

机构 * University of Utah(犹他大学)

Comments 10 pages, 8 figures

Journal ref In Proceedings of the IEEE/CVF International Conference on Computer Vision Workshops (ICCVW), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18564 2025-07-10 cs.CV

Leveraging Local Patch Alignment to Seam-cutting for Large Parallax Image Stitching

Tianli Liao, Chenyang Zhao, Lei Li, Heling Cao

机构 * College of Information Science and Engineering, Henan University of Technology, China(信息科学与工程学院,河南理工大学,中国)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06224 2025-07-09 cs.RO cs.AI

EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow

Yixiang Chen, Peiyan Li, Yan Huang, Jiabing Yang, Kehan Chen, Liang Wang

机构 * New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(模式识别新实验室(NLPR),自动化研究所,中国科学院) School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06072 2025-07-09 cs.CV

MCAM: Multimodal Causal Analysis Model for Ego-Vehicle-Level Driving Video Understanding

Tongtong Cheng, Rongzhen Li, Yixin Xiong, Tao Zhang, Jing Wang, Kai Liu

机构 * Department of Computer Science, Chongqing University, China(重庆大学计算机科学系) National Elite Institute of Engineering, Chongqing University, China(重庆大学工程精英研究院) College of Computer Science and Technology, National University of Deffense Technology, China(国防科技大学计算机科学与技术学院)

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05798 2025-07-09 cs.CV

SPADE: Spatial-Aware Denoising Network for Open-vocabulary Panoptic Scene Graph Generation with Long- and Local-range Context Reasoning

Xin Hu, Ke Qin, Guiduo Duan, Ming Li, Yuan-Fang Li, Tao He

机构 * The Laboratory of Intelligent Collaborative Computing of UESTC(UESTC智能协同计算实验室) Ubiquitous Intelligence and Trusted Services Key Laboratory of Sichuan Province(四川省 Ubiquitous Intelligence and Trusted Services 重点实验室) Monash University(墨尔本大学) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济实验室(深圳))

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05601 2025-07-09 cs.CV

Rethinking Layered Graphic Design Generation with a Top-Down Approach

Jingye Chen, Zhaowen Wang, Nanxuan Zhao, Li Zhang, Difan Liu, Jimei Yang, Qifeng Chen

机构 * HKUST(香港科技大学) Adobe Research(Adobe研究院) Runway

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05536 2025-07-09 cs.CV cs.ET cs.LG

Simulating Refractive Distortions and Weather-Induced Artifacts for Resource-Constrained Autonomous Perception

Moseli Mots'oehli, Feimei Chen, Hok Wai Chan, Itumeleng Tlali, Thulani Babeli, Kyungim Baek, Huaijin Chen

机构 * University of Hawai‘i at Mānoa(夏威夷大学毛伊分校) MindForge AI

Comments This paper has been submitted to the ICCV 2025 Workshop on Computer Vision for Developing Countries (CV4DC) for review

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04801 2025-07-09 cs.CV

PointGAC: Geometric-Aware Codebook for Masked Point Cloud Modeling

Abiao Li, Chenlei Lv, Yuming Fang, Yifan Zuo, Jian Zhang, Guofeng Mei

机构 * Jiangxi University of Finance and Economics(江西财经大学) Shenzhen University(深圳大学) University of Technology Sydney(悉尼科技大学) Fondazione Bruno Kessler(布鲁诺·凯斯勒基金会)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04667 2025-07-09 cs.CV cs.AI cs.MM cs.SD eess.AS

What's Making That Sound Right Now? Video-centric Audio-Visual Localization

Hahyeon Choi, Junhoo Lee, Nojun Kwak

机构 * Seoul National University(首尔国立大学)

Comments Published at ICCV 2025. Project page: https://hahyeon610.github.io/Video-centric_Audio_Visual_Localization/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02395 2025-07-09 cs.CV

Continual Multiple Instance Learning with Enhanced Localization for Histopathological Whole Slide Image Analysis

Byung Hyun Lee, Wongi Jeong, Woojae Han, Kyoungbun Lee, Se Young Chun

机构 * Department of ECE, 2 IPAI, 3 INMC, Seoul National University 4 Department of Pathology, College of Medicine, Seoul National University(1 电子工程系,2 IPAI,3 INMC,首尔国立大学 4 病理学系,医学院,首尔国立大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22099 2025-07-09 cs.CV

BézierGS: Dynamic Urban Scene Reconstruction with Bézier Curve Gaussian Splatting

Zipei Ma, Junzhe Jiang, Yurui Chen, Li Zhang

机构 * School of Data Science, Fudan University(复旦大学数据科学学院)

Comments Accepted at ICCV 2025, Project Page: https://github.com/fudan-zvg/BezierGS

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17695 2025-07-09 cs.CV

MotionDiff: Training-free Zero-shot Interactive Motion Editing via Flow-assisted Multi-view Diffusion

Yikun Ma, Yiqing Li, Jiawei Wu, Xing Luo, Zhi Jin

机构 * Sun Yat-sen University(中山大学) Department of Frontier Research, Peng Cheng Laboratory(前沿研究部,鹏城实验室)

Journal ref ICCV, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01717 2025-07-09 cs.CV

Driving View Synthesis on Free-form Trajectories with Generative Prior

Zeyu Yang, Zijie Pan, Yuankun Yang, Xiatian Zhu, Li Zhang

机构 * Fudan University(复旦大学) University of Surrey(Surrey大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05260 2025-07-08 cs.CV cs.LG cs.RO

Beyond One Shot, Beyond One Perspective: Cross-View and Long-Horizon Distillation for Better LiDAR Representations

Xiang Xu, Lingdong Kong, Song Wang, Chuanwei Zhou, Qingshan Liu

Comments ICCV 2025; 26 pages, 12 figures, 10 tables; Code at http://github.com/Xiangxu-0103/LiMA

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04947 2025-07-08 cs.CV cs.AI

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer

Yecheng Wu, Junyu Chen, Zhuoyang Zhang, Enze Xie, Jincheng Yu, Junsong Chen, Jinyi Hu, Yao Lu, Song Han, Han Cai

机构 * NVIDIA

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23347 2025-07-08 cs.CV

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation

Yi Liu, Shengqian Li, Zuzeng Lin, Feng Wang, Si Liu

机构 * Beihang University(北京航空航天大学) University of Chinese Academy of Sciences(中国科学院大学) Tianjin University(天津大学) CreateAI

Comments Accepted to ICCV 2025. Code available at: https://github.com/IamCreateAI/CycleVAR

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05040 2025-07-08 cs.CV

GaussRender: Learning 3D Occupancy with Gaussian Rendering

Loïck Chambon, Eloi Zablocki, Alexandre Boulch, Mickaël Chen, Matthieu Cord

机构 * ValeoAI Sorbonne University(索邦大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08553 2025-07-08 cs.CV

DynamicFace: High-Quality and Consistent Face Swapping for Image and Video using Composable 3D Facial Priors

Runqi Wang, Yang Chen, Sijie Xu, Tianyao He, Wei Zhu, Dejia Song, Nemo Chen, Xu Tang, Yao Hu

机构 * Xiaohongshu(小红书) ShanghaiTech University(上海科技大学) Shanghai Jiao Tong University(上海交通大学)

Comments Accepted by ICCV 2025. Project page: https://dynamic-face.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15098 2025-07-08 cs.CV cs.AI cs.LG

OminiControl: Minimal and Universal Control for Diffusion Transformer

Zhenxiong Tan, Songhua Liu, Xingyi Yang, Qiaochu Xue, Xinchao Wang

机构 * National University of Singapore(新加坡国立大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04345 2025-07-08 cs.CV

Active Stereo in the Wild through Virtual Pattern Projection

Luca Bartolomei, Matteo Poggi, Fabio Tosi, Andrea Conti, Stefano Mattoccia

Comments IJCV extended version of ICCV 2023 paper: "Active Stereo Without Pattern Projector"

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04685 2025-07-08 cs.CV

TeethGenerator: A two-stage framework for paired pre- and post-orthodontic 3D dental data generation

Changsong Lei, Yaqian Liang, Shaofeng Wang, Jiajia Dai, Yong-Jin Liu

机构 * Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Beijing Stomatological Hospital, Capital Medical University(首都医科大学北京 stomatological Hospital)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏