arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2311.18828 2024-10-08 cs.CV

One-step Diffusion with Distribution Matching Distillation

Tianwei Yin, Michaël Gharbi, Richard Zhang, Eli Shechtman, Fredo Durand, William T. Freeman, Taesung Park

Comments CVPR 2024, Project page: https://tianweiy.github.io/dmd/

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.01484 2024-10-07 cs.CV

Mapping Degeneration Meets Label Evolution: Learning Infrared Small Target Detection with Single Point Supervision

Xinyi Ying, Li Liu, Yingqian Wang, Ruojing Li, Nuo Chen, Zaiping Lin, Weidong Sheng, Shilin Zhou

Journal ref CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01261 2024-10-03 cs.CV

OCC-MLLM:Empowering Multimodal Large Language Model For the Understanding of Occluded Objects

Wenmo Qiu, Xinhan Di

Comments Accepted by CVPR 2024 T4V Workshop (5 pages, 3 figures, 2 tables)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.10229 2024-10-02 cs.CV

OMG-Seg: Is One Model Good Enough For All Segmentation?

Xiangtai Li, Haobo Yuan, Wei Li, Henghui Ding, Size Wu, Wenwei Zhang, Yining Li, Kai Chen, Chen Change Loy

Comments CVPR-2024. Project Page: https://lxtgh.github.io/project/omg_seg/

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.10226 2024-10-02 cs.CV

Towards Language-Driven Video Inpainting via Multimodal Large Language Models

Jianzong Wu, Xiangtai Li, Chenyang Si, Shangchen Zhou, Jingkang Yang, Jiangning Zhang, Yining Li, Kai Chen, Yunhai Tong, Ziwei Liu, Chen Change Loy

Comments CVPR-2024. Project Page: https://jianzongwu.github.io/projects/rovi

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.02332 2024-09-30 cs.CV cs.AI

Efficient Exploration of Image Classifier Failures with Bayesian Optimization and Text-to-Image Models

Adrien LeCoz, Houssem Ouertatani, Stéphane Herbin, Faouzi Adjed

Journal ref Generative Models for Computer Vision - CVPR 2024 Workshop, Jun 2024, Seattle, United States

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00094 2024-09-30 cs.CV cs.AI

Fast ODE-based Sampling for Diffusion Models in Around 5 Steps

Zhenyu Zhou, Defang Chen, Can Wang, Chun Chen

Comments Accepted by CVPR 2024 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.06590 2024-09-30 cs.CV

High-Fidelity GAN Inversion for Image Attribute Editing

Tengfei Wang, Yong Zhang, Yanbo Fan, Jue Wang, Qifeng Chen

Comments CVPR 2022; Project Page is at https://tengfei-wang.github.io/HFGI/

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17759 2024-09-27 eess.IV cs.CV

LGFN: Lightweight Light Field Image Super-Resolution using Local Convolution Modulation and Global Attention Feature Extraction

Zhongxin Yu, Liang Chen, Zhiyun Zeng, Kunping Yang, Shaofei Luo, Shaorui Chen, Cheng Zhong

Comments 10 pages, 5 figures

Journal ref CVPR 2024 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07937 2024-09-27 cs.CV

BOTH2Hands: Inferring 3D Hands from Both Text Prompts and Body Dynamics

Wenqian Zhang, Molin Huang, Yuxuan Zhou, Juze Zhang, Jingyi Yu, Jingya Wang, Lan Xu

Comments Accepted to CVPR 2024

Journal ref Proc. IEEE/CVF Conf. Comput. Vis. Pattern Recognit. (CVPR), 2024, pp. 2393-2404.

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18259 2024-09-27 cs.CV cs.AI

Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Kristen Grauman, Andrew Westbury, Lorenzo Torresani, Kris Kitani, Jitendra Malik, Triantafyllos Afouras, Kumar Ashutosh, Vijay Baiyya, Siddhant Bansal, Bikram Boote, Eugene Byrne, Zach Chavis, Joya Chen, Feng Cheng, Fu-Jen Chu, Sean Crane, Avijit Dasgupta, Jing Dong, Maria Escobar, Cristhian Forigua, Abrham Gebreselasie, Sanjay Haresh, Jing Huang, Md Mohaiminul Islam, Suyog Jain, Rawal Khirodkar, Devansh Kukreja, Kevin J Liang, Jia-Wei Liu, Sagnik Majumder, Yongsen Mao, Miguel Martin, Effrosyni Mavroudi, Tushar Nagarajan, Francesco Ragusa, Santhosh Kumar Ramakrishnan, Luigi Seminara, Arjun Somayazulu, Yale Song, Shan Su, Zihui Xue, Edward Zhang, Jinxu Zhang, Angela Castillo, Changan Chen, Xinzhu Fu, Ryosuke Furuta, Cristina Gonzalez, Prince Gupta, Jiabo Hu, Yifei Huang, Yiming Huang, Weslie Khoo, Anush Kumar, Robert Kuo, Sach Lakhavani, Miao Liu, Mi Luo, Zhengyi Luo, Brighid Meredith, Austin Miller, Oluwatumininu Oguntola, Xiaqing Pan, Penny Peng, Shraman Pramanick, Merey Ramazanova, Fiona Ryan, Wei Shan, Kiran Somasundaram, Chenan Song, Audrey Southerland, Masatoshi Tateno, Huiyu Wang, Yuchen Wang, Takuma Yagi, Mingfei Yan, Xitong Yang, Zecheng Yu, Shengxin Cindy Zha, Chen Zhao, Ziwei Zhao, Zhifan Zhu, Jeff Zhuo, Pablo Arbelaez, Gedas Bertasius, David Crandall, Dima Damen, Jakob Engel, Giovanni Maria Farinella, Antonino Furnari, Bernard Ghanem, Judy Hoffman, C. V. Jawahar, Richard Newcombe, Hyun Soo Park, James M. Rehg, Yoichi Sato, Manolis Savva, Jianbo Shi, Mike Zheng Shou, Michael Wray

Comments Expanded manuscript (compared to arxiv v1 from Nov 2023 and CVPR 2024 paper from June 2024) for more comprehensive dataset and benchmark presentation, plus new results on v2 data release

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16504 2024-09-26 cs.CV

Low Latency Point Cloud Rendering with Learned Splatting

Yueyu Hu, Ran Gong, Qi Sun, Yao Wang

Comments Published at CVPR 2024 Workshop on AIS: Vision, Graphics and AI for Streaming (https://ai4streaming-workshop.github.io/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.13888 2024-09-24 cs.CV cs.LG

Neural Implicit Morphing of Face Images

Guilherme Schardong, Tiago Novello, Hallison Paz, Iurii Medvedev, Vinícius da Silva, Luiz Velho, Nuno Gonçalves

Comments 14 pages, 20 figures, accepted for CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05817 2024-09-24 cs.CV

SAFDNet: A Simple and Effective Network for Fully Sparse 3D Object Detection

Gang Zhang, Junnan Chen, Guohuan Gao, Jianmin Li, Si Liu, Xiaolin Hu

Comments Accepted by CVPR 2024 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.17166 2024-09-23 cs.CV stat.ML

Deep Single Image Camera Calibration by Heatmap Regression to Recover Fisheye Images Under Manhattan World Assumption

Nobuhiko Wakai, Satoshi Sato, Yasunori Ishii, Takayoshi Yamashita

Comments Accepted by CVPR2024

Journal ref 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA, 2024, pp. 11884-11894

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.16193 2024-09-23 cs.CV cs.AI cs.LG cs.MM eess.IV

Improving Multi-label Recognition using Class Co-Occurrence Probabilities

Samyak Rawlekar, Shubhang Bhatnagar, Vishnuvardhan Pogunulu Srinivasulu, Narendra Ahuja

Comments Accepted to ICPR 2024, CVPR workshops 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00842 2024-09-20 cs.CV

An N-Point Linear Solver for Line and Motion Estimation with Event Cameras

Ling Gao, Daniel Gehrig, Hang Su, Davide Scaramuzza, Laurent Kneip

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11347 2024-09-18 cs.AI

Multimodal Datasets and Benchmarks for Reasoning about Dynamic Spatio-Temporality in Everyday Environments

Takanori Ugai, Kensho Hara, Shusaku Egami, Ken Fukuda

Comments 5 pages, 1 figure, 1 table, accepted in Embodied AI 2024 Workshop held in conjunction with CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10603 2024-09-18 cs.CV

CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences

Seungwook Kim, Kejie Li, Xueqing Deng, Yichun Shi, Minsu Cho, Peng Wang

Comments 25 pages, 22 figures, accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.05930 2024-09-18 cs.CV

MED-VT++: Unifying Multimodal Learning with a Multiscale Encoder-Decoder Video Transformer

Rezaul Karim, He Zhao, Richard P. Wildes, Mennatullah Siam

Comments Extension of CVPR'23 paper for journal submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01062 2024-09-17 cs.CV

Layout Agnostic Scene Text Image Synthesis with Diffusion Models

Qilong Zhangli, Jindong Jiang, Di Liu, Licheng Yu, Xiaoliang Dai, Ankit Ramchandani, Guan Pang, Dimitris N. Metaxas, Praveen Krishnan

Comments Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024, pp. 7496-7506

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024, pp. 7496-7506

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.02315 2024-09-11 cs.CV

TempSAL -- Uncovering Temporal Information for Deep Saliency Prediction

Bahar Aydemir, Ludo Hoffstetter, Tong Zhang, Mathieu Salzmann, Sabine Süsstrunk

Comments 10 pages, 7 figures, published in CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00739 2024-09-11 cs.CV

Adversarial Score Distillation: When score distillation meets GAN

Min Wei, Jingkai Zhou, Junyao Sun, Xuesong Zhang

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.02599 2024-09-05 cs.IR cs.CV cs.LG

A Fashion Item Recommendation Model in Hyperbolic Space

Ryotaro Shimizu, Yu Wang, Masanari Kimura, Yuki Hirakawa, Takashi Wada, Yuki Saito, Julian McAuley

Comments This work was presented at the CVFAD Workshop at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10904 2024-09-05 cs.CV

Multi-Task Multi-Modal Self-Supervised Learning for Facial Expression Recognition

Marah Halawa, Florian Blume, Pia Bideau, Martin Maier, Rasha Abdel Rahman, Olaf Hellwich

Comments The paper will appear in the CVPR 2024 workshops proceedings

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2024, pp. 4604-4614

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.00707 2024-09-04 cs.CV cs.AI cs.LG

ReMOVE: A Reference-free Metric for Object Erasure

Aditya Chandrasekar, Goirik Chakrabarty, Jai Bardhan, Ramya Hebbalaguppe, Prathosh AP

Comments Accepted at The First Workshop on the Evaluation of Generative Foundation Models (EvGENFM) at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.13375 2024-09-04 cs.CV

Streamlined Global and Local Features Combinator (SGLC) for High Resolution Image Dehazing

Bilel Benjdira, Anas M. Ali, Anis Koubaa

Comments Accepted in CVPR 2023 Workshops

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, pp. 1855-1864, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06978 2024-09-02 cs.CV

Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

Zhenxin Li, Kailin Li, Shihao Wang, Shiyi Lan, Zhiding Yu, Yishen Ji, Zhiqi Li, Ziyue Zhu, Jan Kautz, Zuxuan Wu, Yu-Gang Jiang, Jose M. Alvarez

Comments The 1st place solution of End-to-end Driving at Scale at the CVPR 2024 Autonomous Grand Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.06267 2024-08-29 cs.CV cs.AI cs.LG cs.SD eess.AS

Multimodality Helps Unimodality: Cross-Modal Few-Shot Learning with Multimodal Models

Zhiqiu Lin, Samuel Yu, Zhiyi Kuang, Deepak Pathak, Deva Ramanan

Comments Published at CVPR 2023. Project site: https://linzhiqiu.github.io/papers/cross_modal/

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.06607 2024-08-27 cs.CV cs.AI cs.CL

Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

Zhang Li, Biao Yang, Qiang Liu, Zhiyin Ma, Shuo Zhang, Jingxu Yang, Yabo Sun, Yuliang Liu, Xiang Bai

Comments CVPR 2024 Highlight

详情

展开后加载摘要…

URL PDF HTML 收藏