arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2404.07155 2024-04-11 cs.CV

Unified Language-driven Zero-shot Domain Adaptation

Senqiao Yang, Zhuotao Tian, Li Jiang, Jiaya Jia

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06918 2024-04-11 cs.CV

HRVDA: High-Resolution Visual Document Assistant

Chaohu Liu, Kun Yin, Haoyu Cao, Xinghua Jiang, Xin Li, Yinsong Liu, Deqiang Jiang, Xing Sun, Linli Xu

Comments Accepted to CVPR 2024 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06832 2024-04-11 cs.CV cs.LG

SplatPose & Detect: Pose-Agnostic 3D Anomaly Detection

Mathis Kruse, Marco Rudolph, Dominik Woiwode, Bodo Rosenhahn

Comments Visual Anomaly and Novelty Detection 2.0 Workshop at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06692 2024-04-11 cs.CV

Perception-Oriented Video Frame Interpolation via Asymmetric Blending

Guangyang Wu, Xin Tao, Changlin Li, Wenyi Wang, Xiaohong Liu, Qingqing Zheng

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.05215 2024-04-11 cs.CV

Spatio-Temporal Attention and Gaussian Processes for Personalized Video Gaze Estimation

Swati Jindal, Mohit Yadav, Roberto Manduchi

Comments Accepted at CVPR 2024 Gaze workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02233 2024-04-11 cs.CV

Visual Concept Connectome (VCC): Open World Concept Discovery and their Interlayer Connections in Deep Models

Matthew Kowal, Richard P. Wildes, Konstantinos G. Derpanis

Comments CVPR 2024 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.10831 2024-04-11 cs.CV cs.AI cs.LG cs.RO

Understanding Video Transformers via Universal Concept Discovery

Matthew Kowal, Achal Dave, Rares Ambrus, Adrien Gaidon, Konstantinos G. Derpanis, Pavel Tokmakov

Comments CVPR 2024 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.04350 2024-04-11 cs.CV

Pre-trained Model Guided Fine-Tuning for Zero-Shot Adversarial Robustness

Sibo Wang, Jie Zhang, Zheng Yuan, Shiguang Shan

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13980 2024-04-11 cs.CV cs.LG

Carve3D: Improving Multi-view Reconstruction Consistency for Diffusion Models with RL Finetuning

Desai Xie, Jiahao Li, Hao Tan, Xin Sun, Zhixin Shu, Yi Zhou, Sai Bi, Sören Pirk, Arie E. Kaufman

Comments 22 pages, 16 figures. Our code, training and testing data, and video results are available at: https://desaixie.github.io/carve-3d. This paper has been accepted to CVPR 2024. v2: incorporated changes from the CVPR 2024 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10908 2024-04-11 cs.CV

CLOVA: A Closed-Loop Visual Assistant with Tool Usage and Update

Zhi Gao, Yuntao Du, Xintong Zhang, Xiaojian Ma, Wenjuan Han, Song-Chun Zhu, Qing Li

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10144 2024-04-11 cs.LG cs.AI cs.CV

Data-Efficient Multimodal Fusion on a Single GPU

Noël Vouitsis, Zhaoyan Liu, Satya Krishna Gorti, Valentin Villecroze, Jesse C. Cresswell, Guangwei Yu, Gabriel Loaiza-Ganem, Maksims Volkovs

Comments CVPR 2024 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00825 2024-04-11 cs.CV cs.AI

SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples

Phillip Howard, Avinash Madasu, Tiep Le, Gustavo Lujan Moreno, Anahita Bhiwandiwalla, Vasudev Lal

Comments Accepted to CVPR 2024. arXiv admin note: text overlap with arXiv:2310.02988

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06622 2024-04-11 cs.CV

Calibrating Higher-Order Statistics for Few-Shot Class-Incremental Learning with Pre-trained Vision Transformers

Dipam Goswami, Bartłomiej Twardowski, Joost van de Weijer

Comments Accepted at CLVision workshop (CVPR 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06542 2024-04-11 cs.CV

Training-Free Open-Vocabulary Segmentation with Offline Diffusion-Augmented Prototype Generation

Luca Barsellotti, Roberto Amoroso, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

Comments CVPR 2024. Project page: https://aimagelab.github.io/freeda/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06337 2024-04-10 cs.CV

Matching 2D Images in 3D: Metric Relative Pose from Metric Correspondences

Axel Barroso-Laguna, Sowmya Munukutla, Victor Adrian Prisacariu, Eric Brachmann

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06065 2024-04-10 cs.CV

Unified Entropy Optimization for Open-Set Test-Time Adaptation

Zhengqing Gao, Xu-Yao Zhang, Cheng-Lin Liu

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06044 2024-04-10 cs.CV

Object Dynamics Modeling with Hierarchical Point Cloud-based Representations

Chanho Kim, Li Fuxin

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.05758 2024-04-10 physics.data-an cs.AI cs.CV cs.LG physics.ao-ph stat.AP

Implicit Assimilation of Sparse In Situ Data for Dense & Global Storm Surge Forecasting

Patrick Ebel, Brandon Victor, Peter Naylor, Gabriele Meoni, Federico Serva, Rochelle Schneider

Comments Accepted at CVPR EarthVision 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.05559 2024-04-10 cs.CV

TIM: A Time Interval Machine for Audio-Visual Action Recognition

Jacob Chalk, Jaesung Huh, Evangelos Kazakos, Andrew Zisserman, Dima Damen

Comments Accepted to CVPR 2024. Project Webpage: https://jacobchalk.github.io/TIM-Project

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00915 2024-04-10 cs.CV cs.RO

Scalable 3D Registration via Truncated Entry-wise Absolute Residuals

Tianyu Huang, Liangzu Peng, René Vidal, Yun-Hui Liu

Comments 24 pages, 12 figures. Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.04932 2024-04-10 cs.CV

Divide and Conquer: High-Resolution Industrial Anomaly Detection via Memory Efficient Tiled Ensemble

Blaž Rolih, Dick Ameln, Ashwin Vaidya, Samet Akcay

Comments To appear at CVPR 24 Visual Anomaly Detection Workshop. Research conducted during Google Summer of Code 2023 at OpenVINO (Intel). GSoC 2023 page: https://summerofcode.withgoogle.com/archive/2023/projects/WUSjdxGl

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03662 2024-04-10 cs.CV

Harnessing Meta-Learning for Improving Full-Frame Video Stabilization

Muhammad Kashif Ali, Eun Woo Im, Dongjin Kim, Tae Hyun Kim

Comments CVPR 2024, Code will be made availble on: http://github.com/MKashifAli/MetaVideoStab

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18078 2024-04-10 cs.CV

Coarse-to-Fine Latent Diffusion for Pose-Guided Person Image Synthesis

Yanzuo Lu, Manlin Zhang, Andy J Ma, Xiaohua Xie, Jian-Huang Lai

Comments Accepted by CVPR 2024 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10634 2024-04-10 cs.CV cs.LG

Anomaly Score: Evaluating Generative Models and Individual Generated Images based on Complexity and Vulnerability

Jaehui Hwang, Junghyuk Lee, Jong-Seok Lee

Comments Accepted in CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10240 2024-04-10 cs.CV

Rich Human Feedback for Text-to-Image Generation

Youwei Liang, Junfeng He, Gang Li, Peizhao Li, Arseniy Klimovskiy, Nicholas Carolan, Jiao Sun, Jordi Pont-Tuset, Sarah Young, Feng Yang, Junjie Ke, Krishnamurthy Dj Dvijotham, Katie Collins, Yiwen Luo, Yang Li, Kai J Kohlhoff, Deepak Ramachandran, Vidhya Navalpakkam

Comments CVPR'24

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09168 2024-04-10 cs.CV cs.GR cs.LG

DiffusionLight: Light Probes for Free by Painting a Chrome Ball

Pakkapon Phongthawee, Worameth Chinchuthakun, Nontaphat Sinsunthithet, Amit Raj, Varun Jampani, Pramook Khungurn, Supasorn Suwajanakorn

Comments CVPR 2024 Oral. For more information and code, please visit our website https://diffusionlight.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02813 2024-04-10 cs.CV cs.AI

BIVDiff: A Training-Free Framework for General-Purpose Video Synthesis via Bridging Image and Video Diffusion Models

Fengyuan Shi, Jiaxi Gu, Hang Xu, Songcen Xu, Wei Zhang, Limin Wang

Comments Accepted by CVPR 2024. Project page: https://bivdiff.github.io; GitHub repository: https://github.com/MCG-NJU/BIVDiff

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18649 2024-04-10 cs.CV

Simple Semantic-Aided Few-Shot Learning

Hai Zhang, Junzhe Xu, Shanlin Jiang, Zhenan He

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17048 2024-04-10 cs.CV

Zero-shot Referring Expression Comprehension via Structural Similarity Between Images and Captions

Zeyu Han, Fangrui Zhu, Qianru Lao, Huaizu Jiang

Comments CVPR 2024, Code available at https://github.com/Show-han/Zeroshot_REC

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05334 2024-04-10 cs.CV

MultIOD: Rehearsal-free Multihead Incremental Object Detector

Eden Belouadah, Arnaud Dapogny, Kevin Bailly

Comments Accepted at the archival track of the Workshop on Continual Learning in Computer Vision (CVPR 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏