arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2303.11726 2024-07-02 cs.CV

3D Human Mesh Estimation from Virtual Markers

Xiaoxuan Ma, Jiajun Su, Chunyu Wang, Wentao Zhu, Yizhou Wang

Comments CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00503 2024-07-02 cs.CV

Toward a Diffusion-Based Generalist for Dense Vision Tasks

Yue Fan, Yongqin Xian, Xiaohua Zhai, Alexander Kolesnikov, Muhammad Ferjad Naeem, Bernt Schiele, Federico Tombari

Comments Published at CVPR 2024 as a workshop paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.16222 2024-07-01 cs.CV

Step Differences in Instructional Video

Tushar Nagarajan, Lorenzo Torresani

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.19393 2024-06-28 cs.CV

Looking 3D: Anomaly Detection with 2D-3D Alignment

Ankan Bhunia, Changjian Li, Hakan Bilen

Comments Accepted at CVPR'24. Codes & dataset available at https://github.com/VICO-UoE/Looking3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18817 2024-06-28 cs.CV cs.AI

Correspondence-Free Non-Rigid Point Set Registration Using Unsupervised Clustering Analysis

Mingyang Zhao, Jingen Jiang, Lei Ma, Shiqing Xin, Gaofeng Meng, Dong-Ming Yan

Comments [CVPR 2024 Highlight] Project and code at: CVPR24_PointSetReg" target="_blank" rel="noopener">https://github.com/zikai1/CVPR24_PointSetReg

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18540 2024-06-28 cs.CV cs.CR

Fully Exploiting Every Real Sample: SuperPixel Sample Gradient Model Stealing

Yunlong Zhao, Xiaoheng Deng, Yijing Liu, Xinjun Pei, Jiazhi Xia, Wei Chen

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18524 2024-06-27 cs.CV

MultiDiff: Consistent Novel View Synthesis from a Single Image

Norman Müller, Katja Schwarz, Barbara Roessle, Lorenzo Porzi, Samuel Rota Bulò, Matthias Nießner, Peter Kontschieder

Comments Project page: https://sirwyver.github.io/MultiDiff Video: https://youtu.be/zBC4z4qXW_4 - CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07544 2024-06-27 cs.CV cs.AI cs.CL cs.LG

Situational Awareness Matters in 3D Vision Language Reasoning

Yunze Man, Liang-Yan Gui, Yu-Xiong Wang

Comments CVPR 2024. Project Page: https://yunzeman.github.io/situation3d

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.16493 2024-06-27 cs.CV

Commonsense Prototype for Outdoor Unsupervised 3D Object Detection

Hai Wu, Shijia Zhao, Xun Huang, Chenglu Wen, Xin Li, Cheng Wang

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.05746 2024-06-27 cs.CV

Editable Scene Simulation for Autonomous Driving via Collaborative LLM-Agents

Yuxi Wei, Zi Wang, Yifan Lu, Chenxin Xu, Changxing Liu, Hao Zhao, Siheng Chen, Yanfeng Wang

Comments CVPR 2024(Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.02416 2024-06-27 cs.CV cs.AI cs.LG cs.RO

ODIN: A Single Model for 2D and 3D Segmentation

Ayush Jain, Pushkal Katara, Nikolaos Gkanatsios, Adam W. Harley, Gabriel Sarch, Kriti Aggarwal, Vishrav Chaudhary, Katerina Fragkiadaki

Comments Camera Ready (CVPR 2024, Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16711 2024-06-27 cs.CV cs.AI cs.HC cs.LG

LEDITS++: Limitless Image Editing using Text-to-Image Models

Manuel Brack, Felix Friedrich, Katharina Kornmeier, Linoy Tsaban, Patrick Schramowski, Kristian Kersting, Apolinário Passos

Comments Proceedings of the 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) The project page is available at https://leditsplusplus-project.static.hf.space

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09858 2024-06-27 cs.CV

Unsupervised Open-Vocabulary Object Localization in Videos

Ke Fan, Zechen Bai, Tianjun Xiao, Dominik Zietlow, Max Horn, Zixu Zhao, Carl-Johann Simon-Gabriel, Mike Zheng Shou, Francesco Locatello, Bernt Schiele, Thomas Brox, Zheng Zhang, Yanwei Fu, Tong He

Comments Accepted by ICCV 2023; Presented on CVPR 2024 Workshop CORR; Project Page:https://github.com/amazon-science/object-centric-vol

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17876 2024-06-27 cs.CV cs.AI cs.CL cs.LG cs.RO

ET tu, CLIP? Addressing Common Object Errors for Unseen Environments

Ye Won Byun, Cathy Jiao, Shahriar Noroozizadeh, Jimin Sun, Rosa Vitiello

Journal ref Conference on Computer Vision and Pattern Recognition (CVPR 2022) - Embodied AI Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17541 2024-06-26 cs.CV

Principal Component Clustering for Semantic Segmentation in Synthetic Data Generation

Felix Stillger, Frederik Hasecke, Tobias Meisen

Comments This is a technical report for a submission to the CVPR "SyntaGen - Harnessing Generative Models for Synthetic Visual Datasets" workshop challenge. The report is already uploaded to the workshop's homepage https://syntagen.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17309 2024-06-26 cs.CV

Zero-Shot Long-Form Video Understanding through Screenplay

Yongliang Wu, Bozheng Li, Jiawang Cao, Wenbo Zhu, Yi Lu, Weiheng Chi, Chuyun Xie, Haolin Zheng, Ziyue Su, Jay Wu, Xu Yang

Comments Highest Score Award to the CVPR'2024 LOVEU Track 1 Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09819 2024-06-26 cs.CV cs.AI

3D Face Tracking from 2D Video through Iterative Dense UV to Image Flow

Felix Taubner, Prashant Raina, Mathieu Tuli, Eu Wern Teh, Chul Lee, Jinmiao Huang

Comments 22 pages, 25 figures, to be published in CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05005 2024-06-26 cs.CV

DITTO: Dual and Integrated Latent Topologies for Implicit 3D Reconstruction

Jaehyeok Shim, Kyungdon Joo

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03806 2024-06-26 cs.CV cs.GR cs.LG

XCube: Large-Scale 3D Generative Modeling using Sparse Voxel Hierarchies

Xuanchi Ren, Jiahui Huang, Xiaohui Zeng, Ken Museth, Sanja Fidler, Francis Williams

Comments CVPR 2024 Highlight. Code: https://github.com/nv-tlabs/XCube/ Website: https://research.nvidia.com/labs/toronto-ai/xcube/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00690 2024-06-26 cs.CV

Open-vocabulary object 6D pose estimation

Jaime Corsetti, Davide Boscaini, Changjae Oh, Andrea Cavallaro, Fabio Poiesi

Comments Camera ready version (CVPR 2024, poster highlight). New Oryon version: arXiv:2406.16384

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.16455 2024-06-25 cs.AI

Guardrails for avoiding harmful medical product recommendations and off-label promotion in generative AI models

Daniel Lopez-Martinez

Comments CVPR 2024 Responsible Generative AI (ReGenAI) workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02072 2024-06-25 cs.CV cs.LG

EGTR: Extracting Graph from Transformer for Scene Graph Generation

Jinbae Im, JeongYeon Nam, Nokyung Park, Hyungmin Lee, Seunghyun Park

Comments CVPR 2024 (Best paper award candidate)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12834 2024-06-25 cs.CV

GroPrompt: Efficient Grounded Prompting and Adaptation for Referring Video Object Segmentation

Ci-Siang Lin, I-Jieh Liu, Min-Hung Chen, Chien-Yi Wang, Sifei Liu, Yu-Chiang Frank Wang

Comments CVPR Workshop (CVinW) 2024. Project page: https://jack24658735.github.io/groprompt/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.03626 2024-06-25 cs.CV

TokenCompose: Text-to-Image Diffusion with Token-level Supervision

Zirui Wang, Zhizhou Sha, Zheng Ding, Yilin Wang, Zhuowen Tu

Comments CVPR 2024, 21 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14993 2024-06-24 cs.CV cs.CL

Disability Representations: Finding Biases in Automatic Image Generation

Yannis Tevissen

Comments Presented at AVA Workshop of CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09201 2024-06-24 cs.CV

Enhanced Object Detection: A Study on Vast Vocabulary Object Detection Track for V3Det Challenge 2024

Peixi Wu, Bosong Chai, Xuan Nie, Longquan Yan, Zeyu Wang, Qifan Zhou, Boning Wang, Yansong Peng, Hebei Li

Journal ref Second Place in CVPR 2024 Vast Vocabulary Visual Detection Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13787 2024-06-21 cs.RO cs.CV

LIT: Large Language Model Driven Intention Tracking for Proactive Human-Robot Collaboration -- A Robot Sous-Chef Application

Zhe Huang, John Pohovey, Ananya Yammanuru, Katherine Driggs-Campbell

Comments Spotlight Presentation at the 3rd Workshop on Computer Vision in the Wild at CVPR 2024. Also accepted by the 5th Annual Embodied AI Workshop at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09469 2024-06-21 cs.CV cs.LG

Virtually Enriched NYU Depth V2 Dataset for Monocular Depth Estimation: Do We Need Artificial Augmentation?

Dmitry Ignatov, Andrey Ignatov, Radu Timofte

Journal ref Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, pages 6177-6186, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.19050 2024-06-21 cs.LG cs.AI

Detecting Generative Parroting through Overfitting Masked Autoencoders

Saeid Asgari Taghanaki, Joseph Lambourne

Comments Accepted to CVPR 2024, Responsible Generative AI workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.00789 2024-06-21 cs.CV

Retrieval-Augmented Egocentric Video Captioning

Jilan Xu, Yifei Huang, Junlin Hou, Guo Chen, Yuejie Zhang, Rui Feng, Weidi Xie

Comments CVPR 2024. Project page is available at: https://jazzcharles.github.io/Egoinstructor/

详情

展开后加载摘要…

URL PDF HTML 收藏