arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2311.17005 2024-05-24 cs.CV

MVBench: A Comprehensive Multi-modal Video Understanding Benchmark

Kunchang Li, Yali Wang, Yinan He, Yizhuo Li, Yi Wang, Yi Liu, Zun Wang, Jilan Xu, Guo Chen, Ping Luo, Limin Wang, Yu Qiao

Comments CVPR 2024 highlight: updated version with Mistral and better performances for MVBench/NExT-QA/STAR/TVQA/EgoSchema/IntentQA

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.12079 2024-05-24 cs.CV

FreeKD: Knowledge Distillation via Semantic Frequency Prompt

Yuan Zhang, Tao Huang, Jiaming Liu, Tao Jiang, Kuan Cheng, Shanghang Zhang

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13194 2024-05-24 cs.CV

KPConvX: Modernizing Kernel Point Convolution with Kernel Attention

Hugues Thomas, Yao-Hung Hubert Tsai, Timothy D. Barfoot, Jian Zhang

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.15406 2024-05-24 cs.CV cs.AI cs.CL cs.MM

Wiki-LLaVA: Hierarchical Retrieval-Augmented Generation for Multimodal LLMs

Davide Caffagni, Federico Cocchi, Nicholas Moratelli, Sara Sarto, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

Comments CVPR 2024 Workshop on What is Next in Multimodal Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16736 2024-05-24 cs.CV

Creating a Digital Twin of Spinal Surgery: A Proof of Concept

Jonas Hein, Frédéric Giraud, Lilian Calvet, Alexander Schwarz, Nicola Alessandro Cavalcanti, Sergey Prokudin, Mazda Farshad, Siyu Tang, Marc Pollefeys, Fabio Carrillo, Philipp Fürnstahl

Comments Accepted for the DCA in MI Workshop @ CVPR 2024. Project page: https://jonashein.github.io/surgerydigitization/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.12979 2024-05-22 cs.CV

OmniGlue: Generalizable Feature Matching with Foundation Model Guidance

Hanwen Jiang, Arjun Karpur, Bingyi Cao, Qixing Huang, Andre Araujo

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.12978 2024-05-22 cs.CV

Personalized Residuals for Concept-Driven Text-to-Image Generation

Cusuh Ham, Matthew Fisher, James Hays, Nicholas Kolkin, Yuchen Liu, Richard Zhang, Tobias Hinz

Comments CVPR 2024. Project page at https://cusuh.github.io/personalized-residuals

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.12725 2024-05-22 cs.CR cs.CV

Nearest is Not Dearest: Towards Practical Defense against Quantization-conditioned Backdoor Attacks

Boheng Li, Yishuo Cai, Haowei Li, Feng Xue, Zhifeng Li, Yiming Li

Comments Accepted to CVPR 2024. 19 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.12531 2024-05-22 cs.CV cs.LG

CustomText: Customized Textual Image Generation using Diffusion Models

Shubham Paliwal, Arushi Jain, Monika Sharma, Vikram Jamwal, Lovekesh Vig

Comments Accepted by AI for Content Creation (AI4CC) workshop at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11905 2024-05-22 cs.CV

CSTA: CNN-based Spatiotemporal Attention for Video Summarization

Jaewon Son, Jaehun Park, Kwangsu Kim

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10998 2024-05-22 cs.CV

ID-Blau: Image Deblurring by Implicit Diffusion-based reBLurring AUgmentation

Jia-Hao Wu, Fu-Jen Tsai, Yan-Tsung Peng, Chung-Chi Tsai, Chia-Wen Lin, Yen-Yu Lin

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.04965 2024-05-22 cs.CV

Seeing a Rose in Five Thousand Ways

Yunzhi Zhang, Shangzhe Wu, Noah Snavely, Jiajun Wu

Comments CVPR 2023. Project page: https://cs.stanford.edu/~yzzhang/projects/rose/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11913 2024-05-21 cs.CV

Diff-BGM: A Diffusion Model for Video Background Music Generation

Sizhe Li, Yiming Qin, Minghang Zheng, Xin Jin, Yang Liu

Comments Accepted by CVPR 2024(Poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11867 2024-05-21 cs.CV cs.LG cs.RO

Depth Prompting for Sensor-Agnostic Depth Estimation

Jin-Hwi Park, Chanhwi Jeong, Junoh Lee, Hae-Gon Jeon

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11643 2024-05-21 cs.CV cs.LG stat.AP

Morphological Prototyping for Unsupervised Slide Representation Learning in Computational Pathology

Andrew H. Song, Richard J. Chen, Tong Ding, Drew F. K. Williamson, Guillaume Jaume, Faisal Mahmood

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11618 2024-05-21 cs.CV cs.AI

Transcriptomics-guided Slide Representation Learning in Computational Pathology

Guillaume Jaume, Lukas Oldenburg, Anurag Vaidya, Richard J. Chen, Drew F. K. Williamson, Thomas Peeters, Andrew H. Song, Faisal Mahmood

Comments CVPR'24, Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11487 2024-05-21 cs.CV

"Previously on ..." From Recaps to Story Summarization

Aditya Kumar Singh, Dhruv Srivastava, Makarand Tapaswi

Comments CVPR 2024; Project page: https://katha-ai.github.io/projects/recap-story-summ/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11483 2024-05-21 cs.CV

MICap: A Unified Model for Identity-aware Movie Descriptions

Haran Raajesh, Naveen Reddy Desanur, Zeeshan Khan, Makarand Tapaswi

Comments CVPR 2024, Project Page: https://katha-ai.github.io/projects/micap/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.11478 2024-05-21 cs.CV eess.IV

Unsupervised Image Prior via Prompt Learning and CLIP Semantic Guidance for Low-Light Image Enhancement

Igor Morawski, Kai He, Shusil Dangi, Winston H. Hsu

Comments Accepted to CVPR 2024 Workshop NTIRE: New Trends in Image Restoration and Enhancement workshop and Challenges

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06592 2024-05-21 cs.CV cs.AI

Exploiting Style Latent Flows for Generalizing Deepfake Video Detection

Jongwook Choi, Taehoon Kim, Yonghyun Jeong, Seungryul Baek, Jongwon Choi

Comments Preprint version, final version will be available at https://openaccess.thecvf.com The IEEE / CVF Computer Vision and Pattern Recognition Conference (CVPR) (2024) Published by: IEEE & CVF

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.15984 2024-05-21 cs.CV cs.LG

Learning Structure-from-Motion with Graph Attention Networks

Lucas Brynte, José Pedro Iglesias, Carl Olsson, Fredrik Kahl

Comments CVPR camera-ready updates

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01933 2024-05-20 cs.CV

PREGO: online mistake detection in PRocedural EGOcentric videos

Alessandro Flaborea, Guido Maria D'Amely di Melendugno, Leonardo Plini, Luca Scofano, Edoardo De Matteis, Antonino Furnari, Giovanni Maria Farinella, Fabio Galasso

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05086 2024-05-20 cs.CV

UFORecon: Generalizable Sparse-View Surface Reconstruction from Arbitrary and UnFavOrable Sets

Youngju Na, Woo Jae Kim, Kyu Beom Han, Suhyeon Ha, Sung-eui Yoon

Comments accepted at CVPR 2024 project page: https://youngju-na.github.io/uforecon.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17614 2024-05-20 cs.CV

Adapt Before Comparison: A New Perspective on Cross-Domain Few-Shot Segmentation

Jonas Herzog

Comments accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.00627 2024-05-20 cs.CV cs.AI

CapHuman: Capture Your Moments in Parallel Universes

Chao Liang, Fan Ma, Linchao Zhu, Yingying Deng, Yi Yang

Comments Accepted by CVPR 2024. Project page: https://caphuman.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10508 2024-05-20 cs.CV

ART3D: 3D Gaussian Splatting for Text-Guided Artistic Scenes Generation

Pengzhi Li, Chengshuai Tang, Qinxuan Huang, Zhiheng Li

Comments Accepted at CVPR 2024 Workshop on AI3DG

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09713 2024-05-20 cs.CV cs.AI cs.CL

SOK-Bench: A Situated Video Reasoning Benchmark with Aligned Open-World Knowledge

Andong Wang, Bo Wu, Sunli Chen, Zhenfang Chen, Haotian Guan, Wei-Ning Lee, Li Erran Li, Chuang Gan

Comments CVPR

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11678 2024-05-20 cs.CV cs.LG

Exploring 3D-aware Latent Spaces for Efficiently Learning Numerous Scenes

Antoine Schnepf, Karim Kassab, Jean-Yves Franceschi, Laurent Caraffa, Flavian Vasile, Jeremie Mary, Andrew Comport, Valérie Gouet-Brunet

Comments Camera-ready version accepted at 3DMV-CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06668 2024-05-20 cs.LG cs.CV

PeerAiD: Improving Adversarial Distillation from a Specialized Peer Tutor

Jaewon Jung, Hongsun Jang, Jaeyong Song, Jinho Lee

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13250 2024-05-20 cs.CV

Video ReCap: Recursive Captioning of Hour-Long Videos

Md Mohaiminul Islam, Ngan Ho, Xitong Yang, Tushar Nagarajan, Lorenzo Torresani, Gedas Bertasius

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏