arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2402.05008 2024-05-20 cs.CV cs.AI cs.LG

EfficientViT-SAM: Accelerated Segment Anything Model Without Accuracy Loss

Zhuoyang Zhang, Han Cai, Song Han

Comments CVPR 2024 Workshop (Efficient Large Vision Models)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07533 2024-05-20 cs.CV

VILA: On Pre-training for Visual Language Models

Ji Lin, Hongxu Yin, Wei Ping, Yao Lu, Pavlo Molchanov, Andrew Tao, Huizi Mao, Jan Kautz, Mohammad Shoeybi, Song Han

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10286 2024-05-17 cs.CV cs.AI

FFF: Fixing Flawed Foundations in contrastive pre-training results in very strong Vision-Language models

Adrian Bulat, Yassine Ouali, Georgios Tzimiropoulos

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10272 2024-05-17 cs.CV cs.AI cs.SD eess.AS eess.IV

Faces that Speak: Jointly Synthesising Talking Face and Speech from Text

Youngjoon Jang, Ji-Hoon Kim, Junseok Ahn, Doyeop Kwak, Hong-Sun Yang, Yoon-Cheol Ju, Il-Hwan Kim, Byeong-Yeol Kim, Joon Son Chung

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10185 2024-05-17 cs.CV

DiverGen: Improving Instance Segmentation by Learning Wider Data Distribution with More Diverse Generative Data

Chengxiang Fan, Muzhi Zhu, Hao Chen, Yang Liu, Weijia Wu, Huaqi Zhang, Chunhua Shen

Comments Accepted to CVPR 2024, codes are available at \href{this https URL}{https://github.com/aim-uofa/DiverGen}

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10053 2024-05-17 cs.CV

SHiNe: Semantic Hierarchy Nexus for Open-vocabulary Object Detection

Mingxuan Liu, Tyler L. Hayes, Elisa Ricci, Gabriela Csurka, Riccardo Volpi

Comments Accepted as a conference paper (highlight) at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09924 2024-05-17 cs.CV

Infrared Adversarial Car Stickers

Xiaopei Zhu, Yuqiu Liu, Zhanhao Hu, Jianmin Li, Xiaolin Hu

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09879 2024-05-17 cs.CV cs.AI

Generative Unlearning for Any Identity

Juwon Seo, Sung-Hoon Lee, Tae-Young Lee, Seungjun Moon, Gyeong-Moon Park

Comments 15 pages, 17 figures, 10 tables, CVPR 2024 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03744 2024-05-17 cs.CV cs.AI cs.CL cs.LG

Improved Baselines with Visual Instruction Tuning

Haotian Liu, Chunyuan Li, Yuheng Li, Yong Jae Lee

Comments Camera ready, CVPR 2024 (highlight). LLaVA project page: https://llava-vl.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09546 2024-05-16 cs.CV

BEHAVIOR Vision Suite: Customizable Dataset Generation via Simulation

Yunhao Ge, Yihe Tang, Jiashu Xu, Cem Gokmen, Chengshu Li, Wensi Ai, Benjamin Jose Martinez, Arman Aydin, Mona Anvari, Ayush K Chakravarthy, Hong-Xing Yu, Josiah Wong, Sanjana Srivastava, Sharon Lee, Shengxin Zha, Laurent Itti, Yunzhu Li, Roberto Martín-Martín, Miao Liu, Pengchuan Zhang, Ruohan Zhang, Li Fei-Fei, Jiajun Wu

Comments CVPR 2024 (Highlight). Project website: https://behavior-vision-suite.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00252 2024-05-16 eess.IV cs.CV

Learned Scanpaths Aid Blind Panoramic Video Quality Assessment

Kanglong Fan, Wen Wen, Mu Li, Yifan Peng, Kede Ma

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18803 2024-05-16 cs.CV cs.CL cs.LG

BioCLIP: A Vision Foundation Model for the Tree of Life

Samuel Stevens, Jiaman Wu, Matthew J Thompson, Elizabeth G Campolongo, Chan Hee Song, David Edward Carlyn, Li Dong, Wasila M Dahdul, Charles Stewart, Tanya Berger-Wolf, Wei-Lun Chao, Yu Su

Comments CVPR 2024 (oral) camera-ready version; data released

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08815 2024-05-15 cs.CV

Efficient Vision-Language Pre-training by Cluster Masking

Zihao Wei, Zixuan Pan, Andrew Owens

Comments CVPR 2024, Project page: https://zxp46.github.io/cluster-masking/ , Code: https://github.com/Zi-hao-Wei/Efficient-Vision-Language-Pre-training-by-Cluster-Masking

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08794 2024-05-15 cs.CV

Ambiguous Annotations: When is a Pedestrian not a Pedestrian?

Luisa Schwirten, Jannes Scholz, Daniel Kondermann, Janis Keuper

Comments Paper accepted at the CVPR 2024 Vision and Language for Autonomous Driving and Robotics Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08720 2024-05-15 cs.CV

The Lost Melody: Empirical Observations on Text-to-Video Generation From A Storytelling Perspective

Andrew Shin, Yusuke Mori, Kunitake Kaneko

Comments To appear at CVPR 2024 Workshop on AI for Content Creation (AI4CC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08483 2024-05-15 cs.CV cs.AI

RDPN6D: Residual-based Dense Point-wise Network for 6Dof Object Pose Estimation Based on RGB-D Images

Zong-Wei Hong, Yen-Yang Hung, Chu-Song Chen

Comments Accepted by CVPR Workshop DLGC, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08458 2024-05-15 cs.CV

Rethinking Prior Information Generation with CLIP for Few-Shot Segmentation

Jin Wang, Bingfeng Zhang, Jian Pang, Honglong Chen, Weifeng Liu

Comments Accepted by CVPR 2024; The camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08322 2024-05-15 cs.CV

StraightPCF: Straight Point Cloud Filtering

Dasith de Silva Edirimuni, Xuequan Lu, Gang Li, Lei Wei, Antonio Robles-Kelly, Hongdong Li

Comments This paper has been accepted to the IEEE/CVF CVPR Conference, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09474 2024-05-15 cs.CV

TCCT-Net: Two-Stream Network Architecture for Fast and Efficient Engagement Estimation via Behavioral Feature Signals

Alexander Vedernikov, Puneet Kumar, Haoyu Chen, Tapio Seppanen, Xiaobai Li

Comments Accepted for the CVPR 2024 workshop (ABAW)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.04394 2024-05-15 cs.CV

Analyzing Participants' Engagement during Online Meetings Using Unsupervised Remote Photoplethysmography with Behavioral Features

Alexander Vedernikov, Zhaodong Sun, Virpi-Liisa Kykyri, Mikko Pohjola, Miriam Nokia, Xiaobai Li

Comments Accepted for the CVPR 2024 workshop (CVPM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11463 2024-05-15 cs.CV

Siamese Learning with Joint Alignment and Regression for Weakly-Supervised Video Paragraph Grounding

Chaolei Tan, Jianhuang Lai, Wei-Shi Zheng, Jian-Fang Hu

Comments Accepted to CVPR 2024. v2: fix a typo in figure 1

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05950 2024-05-15 cs.CL cs.CV cs.LG cs.MM

Language Models as Black-Box Optimizers for Vision-Language Models

Shihong Liu, Zhiqiu Lin, Samuel Yu, Ryan Lee, Tiffany Ling, Deepak Pathak, Deva Ramanan

Comments Published at CVPR 2024. Project site: https://llm-can-optimize-vlm.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.05503 2024-05-15 cs.CV cs.AI cs.LG

Open-world Instance Segmentation: Top-down Learning with Bottom-up Supervision

Tarun Kalluri, Weiyao Wang, Heng Wang, Manmohan Chandraker, Lorenzo Torresani, Du Tran

Comments L3D-IVU Workshop, CVPR 2024. Project page: https://tarun005.github.io/UDOS

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07991 2024-05-14 cs.RO cs.AI cs.CV cs.LG cs.SY eess.SY

SPIN: Simultaneous Perception, Interaction and Navigation

Shagun Uppal, Ananye Agarwal, Haoyu Xiong, Kenneth Shaw, Deepak Pathak

Comments In CVPR 2024. Website at https://spin-robot.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07933 2024-05-14 cs.CV

Authentic Hand Avatar from a Phone Scan via Universal Hand Model

Gyeongsik Moon, Weipeng Xu, Rohan Joshi, Chenglei Wu, Takaaki Shiratori

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07723 2024-05-14 cs.CV

Coarse or Fine? Recognising Action End States without Labels

Davide Moltisanti, Hakan Bilen, Laura Sevilla-Lara, Frank Keller

Comments The Eleventh Workshop on Fine-Grained Visual Categorization (CVPR 24)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07571 2024-05-14 cs.CV

TattTRN: Template Reconstruction Network for Tattoo Retrieval

Lazaro Janier Gonzalez-Soler, Maciej Salwowski, Christian Rathgeb, Daniel Fischer

Comments Accepted at CVPR Workshop 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08529 2024-05-14 cs.CV cs.GR

GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models

Taoran Yi, Jiemin Fang, Junjie Wang, Guanjun Wu, Lingxi Xie, Xiaopeng Zhang, Wenyu Liu, Qi Tian, Xinggang Wang

Comments CVPR 2024, Project page: https://taoranyi.com/gaussiandreamer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07481 2024-05-14 cs.CV

Text Grouping Adapter: Adapting Pre-trained Text Detector for Layout Analysis

Tianci Bi, Xiaoyi Zhang, Zhizheng Zhang, Wenxuan Xie, Cuiling Lan, Yan Lu, Nanning Zheng

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07407 2024-05-14 cs.CV cs.AI

PitcherNet: Powering the Moneyball Evolution in Baseball Video Analytics

Jerrin Bright, Bavesh Balaji, Yuhao Chen, David A Clausi, John S Zelek

Comments IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW'24)

详情

展开后加载摘要…

URL PDF HTML 收藏