arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2312.08914 2024-12-30 cs.CV

CogAgent: A Visual Language Model for GUI Agents

Wenyi Hong, Weihan Wang, Qingsong Lv, Jiazheng Xu, Wenmeng Yu, Junhui Ji, Yan Wang, Zihan Wang, Yuxuan Zhang, Juanzi Li, Bin Xu, Yuxiao Dong, Ming Ding, Jie Tang

Comments CVPR 2024 (Highlight), 27 pages, 19 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16085 2024-12-23 eess.IV cs.CV

Efficient MedSAMs: Segment Anything in Medical Images on Laptop

Jun Ma, Feifei Li, Sumin Kim, Reza Asakereh, Bao-Hiep Le, Dang-Khoa Nguyen-Vu, Alexander Pfefferle, Muxin Wei, Ruochen Gao, Donghang Lyu, Songxiao Yang, Lennart Purucker, Zdravko Marinov, Marius Staring, Haisheng Lu, Thuy Thanh Dao, Xincheng Ye, Zhi Li, Gianluca Brugnara, Philipp Vollmuth, Martha Foltyn-Dumitru, Jaeyoung Cho, Mustafa Ahmed Mahmutoglu, Martin Bendszus, Irada Pflüger, Aditya Rastogi, Dong Ni, Xin Yang, Guang-Quan Zhou, Kaini Wang, Nicholas Heller, Nikolaos Papanikolopoulos, Christopher Weight, Yubing Tong, Jayaram K Udupa, Cahill J. Patrick, Yaqi Wang, Yifan Zhang, Francisco Contijoch, Elliot McVeigh, Xin Ye, Shucheng He, Robert Haase, Thomas Pinetz, Alexander Radbruch, Inga Krause, Erich Kobler, Jian He, Yucheng Tang, Haichun Yang, Yuankai Huo, Gongning Luo, Kaisar Kushibar, Jandos Amankulov, Dias Toleshbayev, Amangeldi Mukhamejan, Jan Egger, Antonio Pepe, Christina Gsaxner, Gijs Luijten, Shohei Fujita, Tomohiro Kikuchi, Benedikt Wiestler, Jan S. Kirschke, Ezequiel de la Rosa, Federico Bolelli, Luca Lumetti, Costantino Grana, Kunpeng Xie, Guomin Wu, Behrus Puladi, Carlos Martín-Isla, Karim Lekadir, Victor M. Campello, Wei Shao, Wayne Brisbane, Hongxu Jiang, Hao Wei, Wu Yuan, Shuangle Li, Yuyin Zhou, Bo Wang

Comments CVPR 2024 MedSAM on Laptop Competition Summary: https://www.codabench.org/competitions/1847/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.04115 2024-12-23 cs.CR

A Stealthy Wrongdoer: Feature-Oriented Reconstruction Attack against Split Learning

Xiaoyang Xu, Mengda Yang, Wenzhe Yi, Ziang Li, Juan Wang, Hongxin Hu, Yong Zhuang, Yaxin Liu

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11228 2024-12-17 cs.CV cs.AI cs.LG

Uni-AdaFocus: Spatial-temporal Dynamic Computation for Video Recognition

Yulin Wang, Haoji Zhang, Yang Yue, Shiji Song, Chao Deng, Junlan Feng, Gao Huang

Comments Accepted by IEEE TPAMI. Journal version of arXiv:2105.03245 (AdaFocusV1, ICCV 2021 Oral), arXiv:2112.14238 (AdaFocusV2, CVPR 2022), and arXiv:2209.13465 (AdaFocusV3, ECCV 2022). Code and pre-trained models: https://github.com/LeapLabTHU/Uni-AdaFocus

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00603 2024-12-17 cs.CV

Hierarchical Memory for Long Video QA

Yiqin Wang, Haoji Zhang, Yansong Tang, Yong Liu, Jiashi Feng, Jifeng Dai, Xiaojie Jin

Comments Accepted to CVPR 2024 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09602 2024-12-16 cs.CV cs.AI cs.LG cs.RO

Hidden Biases of End-to-End Driving Datasets

Julian Zimmerlin, Jens Beißwenger, Bernhard Jaeger, Andreas Geiger, Kashyap Chitta

Comments Technical report for the CVPR 2024 Workshop on Foundation Models for Autonomous Systems. Runner-up of the track 'CARLA Autonomous Driving Challenge' in the 2024 Autonomous Grand Challenge (https://opendrivelab.com/challenge2024/)

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.03500 2024-12-16 cs.LG stat.ML

Optimizing Rank-based Metrics with Blackbox Differentiation

Michal Rolínek, Vít Musil, Anselm Paulus, Marin Vlastelica, Claudio Michaelis, Georg Martius

Comments CVPR 2020 conference paper (oral). The first two authors contributed equally

Journal ref 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06442 2024-12-13 cs.CV cs.RO

QueSTMaps: Queryable Semantic Topological Maps for 3D Scene Understanding

Yash Mehan, Kumaraditya Gupta, Rohit Jayanti, Anirudh Govil, Sourav Garg, Madhava Krishna

Comments Accepted at 2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) as Oral Presentation. Also presented at the 2nd Workshop on Open-Vocabulary 3D Scene Understanding (OpenSUN3D) at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05322 2024-12-10 eess.IV cs.AI cs.CV

$ρ$-NeRF: Leveraging Attenuation Priors in Neural Radiance Field for 3D Computed Tomography Reconstruction

Li Zhou, Changsheng Fang, Bahareh Morovati, Yongtong Liu, Shuo Han, Yongshun Xu, Hengyong Yu

Comments The paper was submitted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04929 2024-12-10 cs.CV cs.AI cs.LG stat.ML

Continuous Video Process: Modeling Videos as Continuous Multi-Dimensional Processes for Video Prediction

Gaurav Shrivastava, Abhinav Shrivastava

Comments Navigate to the project page https://www.cs.umd.edu/~gauravsh/cvp/supp/website.html for video results. Extended version of published CVPR paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02395 2024-12-04 cs.CV

Who Walks With You Matters: Perceiving Social Interactions with Groups for Pedestrian Trajectory Prediction

Ziqian Zou, Conghao Wong, Beihao Xia, Qinmu Peng, Xinge You

Comments 15 pages, 10 figures, submitted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10099 2024-12-04 cs.CV

KP-RED: Exploiting Semantic Keypoints for Joint 3D Shape Retrieval and Deformation

Ruida Zhang, Chenyangguang Zhang, Yan Di, Fabian Manhardt, Xingyu Liu, Federico Tombari, Xiangyang Ji

Comments Accepted by CVPR 2024. We identified an error in our baseline experiments, re-ran them, and updated the results without impacting the paper's conclusions. We apologize for the oversight and appreciate your understanding

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01118 2024-12-03 cs.CV cs.CR

LoyalDiffusion: A Diffusion Model Guarding Against Data Replication

Chenghao Li, Yuke Zhang, Dake Chen, Jingqi Xu, Peter A. Beerel

Comments 13 pages, 6 figures, Submission to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.04101 2024-12-03 cs.LG cs.AI

Continual Learning in the Presence of Repetition

Hamed Hemati, Lorenzo Pellegrini, Xiaotian Duan, Zixuan Zhao, Fangfang Xia, Marc Masana, Benedikt Tscheschner, Eduardo Veas, Yuxiang Zheng, Shiji Zhao, Shao-Yuan Li, Sheng-Jun Huang, Vincenzo Lomonaco, Gido M. van de Ven

Comments Accepted version, to appear in Neural Networks; Challenge Report of the 4th Workshop on Continual Learning in Computer Vision at CVPR

Journal ref Neural Networks, March 2025: Vol 183, 106920

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03077 2024-12-03 cs.CV

MiKASA: Multi-Key-Anchor & Scene-Aware Transformer for 3D Visual Grounding

Chun-Peng Chang, Shaoxiang Wang, Alain Pagani, Didier Stricker

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.13800 2024-12-03 cs.CV

Aligning Step-by-Step Instructional Diagrams to Video Demonstrations

Jiahao Zhang, Anoop Cherian, Yanbin Liu, Yizhak Ben-Shabat, Cristian Rodriguez, Stephen Gould

Comments Accepted to CVPR'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.12307 2024-12-03 cs.CV cs.AI

Predicting and Enhancing the Fairness of DNNs with the Curvature of Perceptual Manifolds

Yanbiao Ma, Licheng Jiao, Fang Liu, Maoji Wen, Lingling Li, Wenping Ma, Shuyuan Yang, Xu Liu, Puhua Chen

Comments 17pages, Accepted by CVPR 2023, Submitted to TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17027 2024-11-27 cs.CV

D$^2$-World: An Efficient World Model through Decoupled Dynamic Flow

Haiming Zhang, Xu Yan, Ying Xue, Zixuan Guo, Shuguang Cui, Zhen Li, Bingbing Liu

Comments The 2nd Place and Innovation Award Solution of Predictive World Model at the CVPR 2024 Autonomous Grand Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15673 2024-11-26 cs.CV

Semantic Shield: Defending Vision-Language Models Against Backdooring and Poisoning via Fine-grained Knowledge Alignment

Alvi Md Ishmam, Christopher Thomas

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15290 2024-11-26 cs.LG cs.NE

GreenMachine: Automatic Design of Zero-Cost Proxies for Energy-Efficient NAS

Gabriel Cortês, Nuno Lourenço, Penousal Machado

Comments Submitted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18068 2024-11-26 cs.CV

Speech2UnifiedExpressions: Synchronous Synthesis of Co-Speech Affective Face and Body Expressions from Affordable Inputs

Uttaran Bhattacharya, Aniket Bera, Dinesh Manocha

Comments 14 pages, 7 figures, 2 tables

Journal ref In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 1st Workshop on Human Motion Generation, 2024, Seattle, Washington, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08919 2024-11-26 cs.CV

CLIP-BEVFormer: Enhancing Multi-View Image-Based BEV Detector with Ground Truth Flow

Chenbin Pan, Burhaneddin Yaman, Senem Velipasalar, Liu Ren

Comments CVPR 2024

Journal ref CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01693 2024-11-26 cs.CV cs.AI

HanDiffuser: Text-to-Image Generation With Realistic Hand Appearances

Supreeth Narasimhaswamy, Uttaran Bhattacharya, Xiang Chen, Ishita Dasgupta, Saayan Mitra, Minh Hoai

Comments Revisions: 1. Added a link to project page in the abstract, 2. Updated references and related work, 3. Fixed some grammatical errors

Journal ref In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024, Seattle, Washington, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.15483 2024-11-22 cs.CV eess.IV

ST-MFNet: A Spatio-Temporal Multi-Flow Network for Frame Interpolation

Duolikun Danier, Fan Zhang, David Bull

Comments Accepted in CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07244 2024-11-21 cs.CV eess.IV

Time-Efficient Light-Field Acquisition Using Coded Aperture and Events

Shuji Habuchi, Keita Takahashi, Chihiro Tsutake, Toshiaki Fujii, Hajime Nagahara

Comments Accepted to IEEE/CVF Computer Vision and Pattern Recognition Conference (CVPR) 2024

Journal ref 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2204.12089 2024-11-21 eess.IV cs.CV

Acquiring a Dynamic Light Field through a Single-Shot Coded Image

Ryoya Mizuno, Keita Takahashi, Michitaka Yoshida, Chihiro Tsutake, Toshiaki Fujii, Hajime Nagahara

Journal ref 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11363 2024-11-19 cs.CV

GPS-Gaussian+: Generalizable Pixel-wise 3D Gaussian Splatting for Real-Time Human-Scene Rendering from Sparse Views

Boyao Zhou, Shunyuan Zheng, Hanzhang Tu, Ruizhi Shao, Boning Liu, Shengping Zhang, Liqiang Nie, Yebin Liu

Comments Journal extension of CVPR 2024,Project page:https://yaourtb.github.io/GPS-Gaussian+

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.15778 2024-11-19 cs.CV cs.AI

ObjectNLQ @ Ego4D Episodic Memory Challenge 2024

Yisen Feng, Haoyu Zhang, Yuquan Xie, Zaijing Li, Meng Liu, Liqiang Nie

Comments The solution for the Natural Language Query track and Goal Step track at CVPR EgoVis Workshop 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01415 2024-11-19 cs.CV

On the Faithfulness of Vision Transformer Explanations

Junyi Wu, Weitai Kang, Hao Tang, Yuan Hong, Yan Yan

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05239 2024-11-19 cs.CV

SwiftBrush: One-Step Text-to-Image Diffusion Model with Variational Score Distillation

Thuan Hoang Nguyen, Anh Tran

Comments Accepted to CVPR 2024; Github: https://github.com/VinAIResearch/SwiftBrush

详情

展开后加载摘要…

URL PDF HTML 收藏